About
This skill analyzes news coverage by querying GDELT to identify independent reporting origins rather than counting duplicate URLs. It collapses reprints and wire copies into single sources while detecting syndication bursts through publication timing analysis. Use it during claim verification to distinguish between genuinely corroborated stories and widespread reprinting of a single source.
Quick Install
Claude Code
Recommendednpx skills add SerhiiKorniienko/bullshit-detector -a claude-code/plugin add https://github.com/SerhiiKorniienko/bullshit-detectorgit clone https://github.com/SerhiiKorniienko/bullshit-detector.git ~/.claude/skills/coverage-checkCopy and paste this command in Claude Code to install this skill
Documentation
coverage-check
Ten URLs are not ten sources. This turns "lots of outlets reported it" into a number you can defend.
When to reach for it
During claim verification, when a claim looks corroborated by volume — a pile of search results all saying the same thing. That pattern has two very different causes:
- Many newsrooms independently established the fact → genuinely strong evidence
- One press release, wire story, or study got reprinted 40 times → one source
Search results look identical in both cases. This tells them apart.
Usage
uv run scripts/coverage.py "<query>" [--timespan 3m] [--max 250] [--sort dateasc] [--timeout 120] [--json]
It is slow. This is normal. GDELT takes ~15s for a trivial one-day query and considerably longer
for a 3-month window at 250 records. The script prints progress to stderr and how long the call took,
so you can tell "working" from "hung" — if you see the querying line, wait. Narrowing --timespan is
the speed lever; raise --timeout before assuming it's broken.
The query accepts GDELT operators: "exact phrase", (a OR b), -exclude,
domain:example.com, sourcelang:english. Quote the distinctive phrasing of the claim — a
verbatim phrase is what catches reprints.
# Is this "40 outlets confirmed it" or one wire story?
uv run scripts/coverage.py '"quantum breakthrough" AND university'
# Narrow to the week the claim surfaced
uv run scripts/coverage.py '"record quarterly revenue" domain:reuters.com' --timespan 7d
Reading the output
The first line is the verdict the detector needs. The rest supports it.
- Distinct story clusters — articles grouped by headline similarity. This is the origin estimate. Outlets ≫ clusters means syndication.
- ⚠️ syndicated on a cluster — multiple outlets published the same story inside 24h. Treat the whole cluster as one source.
- Span (hours) — a tight burst points at a press release or embargo lift; coverage developed over weeks is more likely independent.
Feed the result into the report's evidence column as an origin count: "6 results, 1 origin (all reprints of the company's press release)" is worth more than six links.
Limits — read these before trusting a number
-
Rolling 3-month window only. GDELT DOC 2.0 does not reach further back. For an older claim this returns nothing, and nothing does not mean unreported. The script says so in its output; don't let the agent quietly read empty as disconfirming.
-
Clustering is headline similarity, on two measures: sequence ratio for reworded headlines and token overlap for the same facts in a different order. Grouping is transitive — three outlets on one wire story stay together even when the two extremes score below the bar individually. Verbatim reprints, rewritten wire copy and reordered headlines all collapse correctly.
What it still won't catch: two newsrooms that independently reached the same finding and described it in genuinely different words. Those show as separate clusters, which is the safe direction to be wrong in — it under-reports syndication rather than inventing it.
Thresholds were tuned against real GDELT output, not guessed. If you see false merges, raise
TITLE_MATCH/TOKEN_MATCHin the script; if wire copy slips through as distinct, lower them. -
Presence is not credibility. A claim covered by 200 outlets in 30 distinct clusters is widely reported, not true. Verdicts still need the source hierarchy in the detector's RUBRIC.md.
-
Results cap at 250 per query. When the cap is hit the output says so — every count becomes a lower bound, and the honest fix is a narrower
--timespan, not a bigger number. -
The free endpoint is unreliable, and this is the important one. GDELT returns "Please limit requests to one every 5 seconds" well below that rate whenever its public API is busy — independent of IP, User-Agent, and query size. Measured behaviour: identical calls succeed and fail minutes apart. The script retries with growing backoff and then exits 3.
Exit 3 means "unmeasured", not "no coverage". Never let a failed check weaken or strengthen a verdict, and never record it as though the search came back empty. If the tool can't measure, the report says the origin count is unknown and falls back to the eyeball tells in RUBRIC.md. Retry in a few minutes, or skip it.
Exit Meaning 0 measurement succeeded (including a legitimate zero-result window) 1 bad input or unreachable host 3 GDELT throttled — no measurement, claim is unmeasured -
No API key, no auth, free. Nothing to configure, nothing to rotate.
What it does not do
It counts and groups coverage. It does not fetch article text — that's fetch-content — and it does not judge anything. Analysis skills read its output; they never call it to decide a verdict on their own.
GitHub Repository
Frequently asked questions
What is the coverage-check skill?
coverage-check is a Claude Skill by SerhiiKorniienko. Skills package instructions and resources that Claude loads on demand, so Claude can perform coverage-check-related tasks without extra prompting.
How do I install coverage-check?
Use the install commands on this page: add coverage-check to Claude Code as a plugin, or clone its repository into your skills directory, then restart Claude so it picks up the skill.
What category does coverage-check belong to?
coverage-check is in the Design category, tagged ai.
Is coverage-check free to use?
Yes. coverage-check is listed on AIMCP and free to install.
Related Skills
Use the executing-plans skill when you have a complete implementation plan to execute in controlled batches with review checkpoints. It loads and critically reviews the plan, then executes tasks in small batches (default 3 tasks) while reporting progress between each batch for architect review. This ensures systematic implementation with built-in quality control checkpoints.
This skill dispatches a code-reviewer subagent to analyze code changes against requirements before proceeding. It should be used after completing tasks, implementing major features, or before merging to main. The review helps catch issues early by comparing the current implementation with the original plan.
This skill provides a comprehensive guide for developers to connect MCP servers to Claude Code using HTTP, stdio, or SSE transports. It covers installation, configuration, authentication, and security for integrating external services like GitHub, Notion, and custom APIs. Use it when setting up MCP integrations, configuring external tools, or working with Claude's Model Context Protocol.
This skill helps developers choose between Claude Code Web and CLI interfaces based on task analysis, then enables seamless session teleportation between these environments. It optimizes workflow by managing session state and context when switching between web, CLI, or mobile. Use it for complex projects requiring different tools at various stages.
