Preview the sources a topic extraction would hit
Runs the exact same source selection as the extraction trigger — the top sources (by citation count) matching the standard analytics filters — but enqueues nothing. Returns the selected sources grouped by domain so the frontend can let the user assess the scope before triggering.
Headers
Project ID to specify the project context
Body
Start date (inclusive)
"2025-01-01"
End date (exclusive)
"2025-02-01"
Filter by country codes
Filter by language codes
Filter by AI models
Filter by query IDs
Filter by source presence: "sources" (only with sources), "no_sources" (only without), "all" (no filter). Legacy true/false values are still accepted.
all, sources, no_sources Filter by shopping presence: "shopping" (only with shopping), "no_shopping" (only without), "all" (no filter). Legacy true/false values are still accepted.
all, shopping, no_shopping Filter by query tag IDs (numeric — bigint column)
Filter by execution tag IDs
Filter by query tag group IDs — matches rows carrying any tag filed under a selected group. Combines with queryTagIds per queryTagMode.
Filter by execution tag group IDs — matches rows carrying any tag filed under a selected group. Combines with execTagIds per execTagMode.
Filter by query type. Include "untyped" to also match queries without a type. Results from since-deleted queries are excluded when this filter is set.
comparative, informative, perception, untyped Query tag matching mode: "or" matches ANY tag (default), "and" matches ALL tags.
and, or Execution tag matching mode: "or" matches ANY tag (default), "and" matches ALL tags.
and, or IANA timezone for date bucketing and filtering (e.g. "Europe/Brussels"). Defaults to UTC.
"Europe/Brussels"
Row grouping level for entity results. "none" = one row per entity, "division" = one row per division (entities not in a division get their own row), "group" = one row per top-level group (divisions roll up into their group). Defaults to "none".
none, division, group Deprecated — use groupBy instead. true is equivalent to groupBy="group".
false
Include entities that have only ever been seen in a single AI response. These are mostly one-off extraction noise and are hidden by default.
false
Response
Number of source snapshots (distinct url + month) selected — the exact set the extraction endpoint would enqueue. Capped at limit.
100
Maximum number of sources selected (top-N by citation count).
100
Estimated analysis coverage of the selected snapshots: withMeta is guaranteed in the topic map, preMeta will be skipped, uncertain is resolved at run time. Buckets sum to totalSources.
Selected sources grouped by domain, ordered by sourceCount descending (then citationCount). Lets the user assess what the extraction will hit.
Source-categorisation coverage of the selected domains — check before extracting whether the sources have been categorised.
Project-level relationship-overlay readiness. Content-gap presence (owned/competitor) is only meaningful once these are mapped.