MCP Server

data-breach-detector

io.github.beepboop2025/data-breach-detector
Data & Analytics Security Public & reachable MCP 2025-11-25

What this MCP does

Searches historical and recent public breach disclosures, checks organizational exposure, aggregates breach statistics, and performs local threat-text assessment.

assess_threat
Classify a piece of security text you supply — an advisory, alert or forum post — into a threat level, matched categories, financial-target flags, a confidence score and a recommended action. Pure local analysis: it collects nothing, stores nothing and reaches no network; the text never leaves the server. Use it to triage findings surfaced by breach_news or from your own monitoring.
Read only Idempotent
Input schema
{'type': 'object', 'title': 'assess_threatArguments', 'required': ['text'], 'properties': {'text': {'type': 'string', 'title': 'Text', 'description': 'the security text to classify â\x80\x94 an advisory, alert or forum post'}}}
breach_history
Search the FULL historical breach archive — every incident this server knows about, back to 2007: HaveIBeenPwned's verified breach directory, the 2020-2025 ransomwatch leak-site archive (~16k victims), the RansomLook live tracker and SEC 8-K Item 1.05 filings. Filter by keyword, year range, sector, exposed data type or minimum scale; order by date or size. Returns disclosure metadata only, never breach contents. Use this for questions like 'what were the biggest breaches of 2013' or 'which airlines have ever been hit by ransomware'.
Read only Open world Idempotent
Input schema
{'type': 'object', 'title': 'breach_historyArguments', 'properties': {'limit': {'type': 'integer', 'title': 'Limit', 'default': 10, 'maximum': 100, 'minimum': 1, 'description': 'maximum incidents to return (default 10; raise it deliberately, large pages are heavy for an agent loop)'}, 'order': {'type': 'string', 'title': 'Order', 'default': 'newest', 'description': "'newest' (default), 'oldest' or 'largest' (by accounts exposed)"}, 'query': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Query', 'default': None, 'description': 'optional keyword over entity, title, summary, actor and data types; omit to browse the whole archive'}, 'offset': {'type': 'integer', 'title': 'Offset', 'default': 0, 'minimum': 0, 'description': 'how many matching incidents to skip before the page starts; count can run to five figures over the ~16k-post archive, so this is how the tail is reached'}, 'sector': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Sector', 'default': None, 'description': "industry keyword filter, e.g. 'bank', 'health', 'gaming'"}, 'year_to': {'anyOf': [{'type': 'integer', 'minimum': 2000}, {'type': 'null'}], 'title': 'Year To', 'default': None, 'description': 'latest incident year to include, e.g. 2020'}, 'data_type': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Data Type', 'default': None, 'description': "require an exposed data type, e.g. 'passwords', 'credit card', 'health'"}, 'year_from': {'anyOf': [{'type': 'integer', 'minimum': 2000}, {'type': 'null'}], 'title': 'Year From', 'default': None, 'description': 'earliest incident year to include, e.g. 2013'}, 'min_accounts': {'type': 'integer', 'title': 'Min Accounts', 'default': 0, 'minimum': 0, 'description': 'only incidents exposing at least this many accounts'}}}
breach_news
Read recent breach and ransomware DISCLOSURES from public threat-intel feeds (HaveIBeenPwned, the RansomLook live leak-site tracker and SEC 8-K Item 1.05 filings), newest first. Every row is metadata only — entity, date, scale, exposed data TYPES, threat level and source — never the leaked data, and a redaction pass strips anything credential-shaped before it is returned. Use sector to narrow to an industry keyword; for one specific organization use check_exposure; for all-time history use breach_history.
Read only Open world Idempotent
Input schema
{'type': 'object', 'title': 'breach_newsArguments', 'properties': {'limit': {'type': 'integer', 'title': 'Limit', 'default': 10, 'maximum': 100, 'minimum': 1, 'description': 'maximum disclosures to return (default 10; raise it deliberately, large pages are heavy for an agent loop)'}, 'offset': {'type': 'integer', 'title': 'Offset', 'default': 0, 'minimum': 0, 'description': 'how many matching disclosures to skip before the page starts; with limit this walks a result set larger than any single page (count reports the full total)'}, 'sector': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Sector', 'default': None, 'description': "optional keyword filter over entity, title, summary, categories and exposed data types, e.g. 'bank', 'health', 'crypto'"}, 'source': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Source', 'default': None, 'description': "optional source filter: 'HaveIBeenPwned', 'RansomLook', 'ransomwatch-archive' or 'SEC EDGAR 8-K 1.05'"}, 'since_days': {'type': 'integer', 'title': 'Since Days', 'default': 30, 'minimum': 1, 'description': 'look-back window in days over disclosure dates (default 30)'}}}
breach_stats
Aggregate the full breach archive into analyst-grade statistics: incidents and accounts exposed per year, per source, per exposed data type, per threat level, or per ransomware actor — plus the five largest incidents ever recorded. Use it to answer 'how has breach volume trended since 2015', 'which ransomware groups have the most victims' or 'how often are passwords part of a breach'. Aggregate counts only; no leaked records.
Read only Open world Idempotent
Input schema
{'type': 'object', 'title': 'breach_statsArguments', 'properties': {'limit': {'type': 'integer', 'title': 'Limit', 'default': 40, 'maximum': 500, 'minimum': 1, 'description': 'how many buckets to return, largest first (default 40); buckets_total reports how many exist, and grouping by actor over the ~16k-post archive produces far more'}, 'sector': {'anyOf': [{'type': 'string'}, {'type': 'null'}], 'title': 'Sector', 'default': None, 'description': 'optional industry keyword filter applied before aggregating'}, 'group_by': {'type': 'string', 'title': 'Group By', 'default': 'year', 'description': "aggregation axis: 'year' (default), 'source', 'data_type', 'threat_level' or 'actor' (ransomware group)"}}}
breach_timeline
Build the incident-by-incident CHRONOLOGY of one organization across every source and all history, with judgment on top: first and latest incident, incidents per year, whether the organization is a repeat victim, worst threat level and total accounts ever exposed. Those summary fields cover EVERY incident on record. The timeline list carries a window of them, oldest first within the window, defaulting to the most recent limit incidents and paging backwards with offset, so an organization with a long history shows its current state first rather than only its ancient one. Repeat victimhood is a forward-looking risk signal: organizations named more than once have demonstrably not closed the gap. Metadata only; never the leaked data. For a yes/no presence check use check_exposure.
Read only Open world Idempotent
Input schema
{'type': 'object', 'title': 'breach_timelineArguments', 'required': ['entity'], 'properties': {'limit': {'type': 'integer', 'title': 'Limit', 'default': 12, 'maximum': 100, 'minimum': 1, 'description': 'how many incidents the timeline list carries (default 12); the counts, span and judgment always cover every incident'}, 'entity': {'type': 'string', 'title': 'Entity', 'description': "domain, company or brand to build the chronology for, e.g. 'yahoo.com' or 'Adobe'"}, 'offset': {'type': 'integer', 'title': 'Offset', 'default': 0, 'minimum': 0, 'description': 'pages backwards through the chronology from the recent end: 0 gives the newest window, 12 gives the window before that'}}}
check_exposure
Answer whether a domain, company or brand appears in public breach or ransomware DISCLOSURES across ALL history (2007 → today): yes/no with mention count, worst threat level, total accounts exposed across matches, the exposed data TYPES, and the matching disclosure metadata — never the exposed records themselves. This is a triage signal built from disclosure feeds, not proof of compromise; confirm through authorized channels before acting. For the incident-by-incident chronology of one entity, use breach_timeline; for a recent-news sweep, use breach_news. mentions, the aggregates and the data types always cover every match; matches carries one page of them, sized by limit and walked with offset.
Read only Open world Idempotent
Input schema
{'type': 'object', 'title': 'check_exposureArguments', 'required': ['query'], 'properties': {'limit': {'type': 'integer', 'title': 'Limit', 'default': 8, 'maximum': 100, 'minimum': 1, 'description': 'maximum matching disclosures to return (default 8; the mention count and the aggregates always cover every match)'}, 'query': {'type': 'string', 'title': 'Query', 'description': "domain, company or brand to look up, e.g. 'example.com' or 'Acme'"}, 'offset': {'type': 'integer', 'title': 'Offset', 'default': 0, 'minimum': 0, 'description': 'how many matches to skip before the page starts; with limit this reaches matches beyond the first page'}, 'since_days': {'type': 'integer', 'title': 'Since Days', 'default': 100000, 'minimum': 1, 'description': 'optional look-back window in days; the default covers all history'}}}
feed_sources
List the public disclosure feeds this server aggregates, how many disclosures are cached per source, each source's newest item and an honest staleness flag, plus cache ages. Takes no arguments. Also states the scope plainly: public feeds only — no .onion access, no arbitrary fetching or crawling, no credential or PII output. Check this first if another tool's answer looks thin: a stale live feed is a finding, not background noise.
Read only Open world Idempotent
Input schema
{'type': 'object', 'title': 'feed_sourcesArguments', 'properties': {}}
Added
feed_sources
Sept. 17, 2026, 12:40 p.m.
Added
assess_threat
Sept. 17, 2026, 12:40 p.m.
Added
breach_stats
Sept. 17, 2026, 12:40 p.m.
Added
breach_timeline
Sept. 17, 2026, 12:40 p.m.
Added
breach_history
Sept. 17, 2026, 12:40 p.m.
Added
check_exposure
Sept. 17, 2026, 12:40 p.m.
Added
breach_news
Sept. 17, 2026, 12:40 p.m.

Crypto Bot Audit + Market Data (x402 paid)

io.github.kaminariouji/x402-audit-agent

Audits crypto bot source code and provides paid crypto market, token, gas, stablecoin, and DeFi TVL data.

TunnelMind Data API

ai.tunnelmind/data

Aggregates web, routing, supply-chain, tracker, threat, and agent-registry intelligence into risk verdicts, evidence, receipts, a…

Polyform

org.polyform/polyform

Offers pay-per-call APIs for economic, financial, geographic, healthcare, domain, email, sanctions, blockchain, and business risk…

HubVibe: Pay-per-Call Tools for AI Agents: Web Search, Email Verify, KYC, Stocks, Crypto, News, Data

io.github.Its-fortunatefolly/hubvibe

Offers paid utilities for web audits, HTTP fetching and extraction, BigQuery analysis, LLM processing, code execution, blockchain…

Qiniso

io.github.qinisolabs/qiniso

Provides deterministic formatting, parsing, holiday and tax lookups, address handling, and checksum or structure validation for i…

SYNTHORA x402 Intelligence Mesh

com.hergertsynthora/synthora-x402

Delivers paid intelligence on markets, crypto, prediction markets, maritime and space risks, domain and wallet exposure, sanction…

Satoshidata Wallet Intel

io.github.wrbtc/wallet-intelligence

Provides Bitcoin wallet intelligence, address labels, trust and risk signals, transaction verification, entity activity, fees, me…

Cloudflare Radar

io.github.pipeworx-io/cloudflare-radar

Provides Cloudflare Radar internet observatory data on DDoS attacks, BGP leaks, domain popularity, internet quality, and traffic …