Talkwalker: Most-used fields#
The table below gives information about most-used fields that you can import from Talkwalker. Other fields might also be available in Adverity.
The fields that you can fetch in Adverity are updated regularly to reflect updates to data source APIs.
API name |
Adverity UI name |
Description |
Use in Adverity |
|---|---|---|---|
DEPRECATED_spam_level |
DEPRECATED_spam_level |
(DEPRECATED) An integer representing the spam level of the document, on a scale of 0 to 100. |
metric |
article_extended_attributes.bluesky_likes |
article_extended_attributes.bluesky_likes |
The number of likes the article received on Bluesky. |
metric |
article_extended_attributes.bluesky_quotes |
article_extended_attributes.bluesky_quotes |
The number of quotes the article received on Bluesky. |
metric |
article_extended_attributes.bluesky_reposts |
article_extended_attributes.bluesky_reposts |
The number of reposts the article received on Bluesky. |
metric |
article_extended_attributes.bluesky_shares |
article_extended_attributes.bluesky_shares |
The number of shares the article received on Bluesky. |
metric |
article_extended_attributes.facebook_likes |
article_extended_attributes.facebook_likes |
The number of likes the article received on Facebook. |
metric |
article_extended_attributes.facebook_reactions_total |
article_extended_attributes.facebook_reactions_total |
The total number of reactions (likes, loves, wows, etc.) the article received on Facebook. |
metric |
article_extended_attributes.facebook_shares |
article_extended_attributes.facebook_shares |
The number of shares the article received on Facebook. |
metric |
article_extended_attributes.num_comments |
article_extended_attributes.num_comments |
The total number of comments associated with the article or post. |
metric |
article_extended_attributes.twitter_shares |
article_extended_attributes.twitter_shares |
The number of shares (retweets) the article received on Twitter. |
metric |
article_extended_attributes.youtube_likes |
article_extended_attributes.youtube_likes |
The number of likes the video received on YouTube. |
metric |
article_extended_attributes.youtube_views |
article_extended_attributes.youtube_views |
The total number of views the video received on YouTube. |
metric |
cluster_id |
cluster_id |
A unique identifier for a cluster of similar documents. |
dimension |
content |
content |
The full textual content of the document. |
dimension |
content_snippet |
content_snippet |
A snippet of the document’s content, often highlighting parts that match the search query. |
dimension |
domain_url |
domain_url |
The domain URL of the document’s source (e.g., example.com), used for filtering across an entire domain. |
dimension |
engagement_rate |
engagement_rate |
metric |
|
entity_url |
entity_url |
A list of URLs for entities (e.g., persons, brands) extracted or linked from the article. |
dimension |
estimated_reach |
estimated_reach |
The estimated number of unique individuals reached by the article or post. |
metric |
extra_author_attributes.description |
extra_author_attributes.description |
The description of the author associated with the document. |
dimension |
extra_author_attributes.gender |
extra_author_attributes.gender |
The gender of the author. |
dimension |
extra_author_attributes.id |
extra_author_attributes.id |
A unique identifier for the author of the document. |
dimension |
extra_author_attributes.image_url |
extra_author_attributes.image_url |
dimension |
|
extra_author_attributes.name |
extra_author_attributes.name |
The name of the author of the document. |
dimension |
extra_author_attributes.short_name |
extra_author_attributes.short_name |
dimension |
|
extra_author_attributes.url |
extra_author_attributes.url |
dimension |
|
extra_source_attributes.description |
extra_source_attributes.description |
The description of the source (e.g., website, publication) of the document. |
dimension |
extra_source_attributes.id |
extra_source_attributes.id |
dimension |
|
extra_source_attributes.name |
extra_source_attributes.name |
dimension |
|
extra_source_attributes.url |
extra_source_attributes.url |
dimension |
|
extra_source_attributes.world_data.city |
extra_source_attributes.world_data.city |
The city where the source of the document is located. |
dimension |
extra_source_attributes.world_data.continent |
extra_source_attributes.world_data.continent |
The continent where the source of the document is located. |
dimension |
extra_source_attributes.world_data.country |
extra_source_attributes.world_data.country |
The country where the source of the document is located. |
dimension |
extra_source_attributes.world_data.country_code |
extra_source_attributes.world_data.country_code |
The ISO 2-character country code where the source of the document is located. |
dimension |
extra_source_attributes.world_data.latitude |
extra_source_attributes.world_data.latitude |
The geographical latitude of the source’s location. |
metric |
extra_source_attributes.world_data.longitude |
extra_source_attributes.world_data.longitude |
The geographical longitude of the source’s location. |
metric |
extra_source_attributes.world_data.region |
extra_source_attributes.world_data.region |
The region within the country where the source of the document is located. |
dimension |
extra_source_attributes.world_data.resolution |
extra_source_attributes.world_data.resolution |
dimension |
|
fakenews_level |
fakenews_level |
An integer indicating the detected fakenews level of the document, on a scale of 0 to 100. |
metric |
fluency_level |
fluency_level |
(DEPRECATED) An integer representing the fluency level of the document’s language, on a scale of 0 to 100. |
metric |
host_url |
host_url |
The host URL of the document’s source (e.g., www.example.com or blog.example.com), used for filtering on specific hosts. |
dimension |
iab_category.0.tier1 |
iab_category.0.tier1 |
dimension |
|
iab_category.1.tier1 |
iab_category.1.tier1 |
dimension |
|
iab_category.2.tier1 |
iab_category.2.tier1 |
dimension |
|
iab_category.2.tier2 |
iab_category.2.tier2 |
dimension |
|
images |
images |
A list of image objects associated with the document, including their URLs, legends, width, and height. |
dimension |
indexed |
indexed |
The timestamp (in milliseconds since epoch) when the document was initially indexed in the Talkwalker system. |
dimension |
lang |
lang |
The 2-character ISO code for the language of the document’s content. |
dimension |
noise_category |
noise_category |
The category of noise detected in the document, such as “promotions”, “hate_speech”, or “job_offers”. |
dimension |
noise_level |
noise_level |
An integer indicating the detected noise level of the document, on a scale of 0 to 100. |
metric |
parent_url |
parent_url |
The URL of the parent document in a conversation thread (e.g., the original post for a comment or retweet). |
dimension |
porn_level |
porn_level |
An integer indicating the detected pornographic content level of the document, on a scale of 0 to 100. |
metric |
post_type |
post_type |
The type of post, such as “TEXT”, “VIDEO”, “LINK”, or “AUDIO”. |
dimension |
published |
published |
The timestamp (in milliseconds since epoch) indicating when the document was published. This field can be used for time-based filtering and histogram breakdowns. |
dimension |
reach |
reach |
The estimated number of unique people who were exposed to or reached by the article or post. |
metric |
report_date |
report_date |
dimension |
|
root_url |
root_url |
The root URL of the document’s source (e.g., https://www.example.com/). |
dimension |
search_indexed |
search_indexed |
The timestamp (in milliseconds since epoch) indicating when the document was indexed by Talkwalker. This field can be used for time-based filtering and histogram breakdowns. |
dimension |
sentiment |
sentiment |
The sentiment score of the document, determined by Talkwalker’s Natural Language Processing (NLP), indicating whether the content is positive, negative, or neutral. |
dimension |
source_extended_attributes.alexa_pageviews |
source_extended_attributes.alexa_pageviews |
The number of page views according to Alexa for the source of the document. |
metric |
source_extended_attributes.alexa_unique_visitors |
source_extended_attributes.alexa_unique_visitors |
The number of unique visitors according to Alexa for the source of the document. |
metric |
source_extended_attributes.bluesky_followers |
source_extended_attributes.bluesky_followers |
The number of followers the source has on Bluesky. |
metric |
source_extended_attributes.linkedin_followers |
source_extended_attributes.linkedin_followers |
metric |
|
source_type |
source_type |
The media type code of the document’s source (e.g., “ONLINENEWS_NEWSPAPER”, “SOCIALMEDIA”). |
dimension |
tags_internal |
tags_internal |
Internal tags applied to the document (e.g., “hasImage”, “isQuestion”). |
dimension |
title |
title |
The title of the document. |
dimension |
title_snippet |
title_snippet |
A snippet of the document’s title, often used for highlighting matching keywords. |
dimension |
url |
url |
The unique URL of the document. This field serves as the primary key. |
dimension |
videos |
videos |
A list of video objects associated with the document, including their URLs, legends, width, and height. |
dimension |
word_count |
word_count |
The total number of words in the document’s content. |
metric |