committee-releases
Enables searching and retrieving documents that congressional committees publish on their own websites, including press releases, oversight letters, staff reports, and interview transcripts, with no authentication required.
README
@pipeworx/committee-releases
Documents congressional committees publish on their own websites — press releases, oversight letters, staff reports, released interview transcripts and investigation files. Keyless.
Tools
search_committee_documents(...)— find documents by headline across seven committees, filtered bycommittee,doc_typeand (where dates exist)since.get_committee_document(...)— a document page's headline, publication date, published text and attached PDFs with URLs and sizes.list_committee_document_types(...)— which committees are covered and what each publishes.
Auth
None.
Why this exists
congressional-documents covers GovInfo — the official printed record, which lags the event by months and never receives the ad-hoc material a committee posts to its own site. This covers that gap, within hours of publication.
Committees covered
| Committee | Chamber | Shape | Dates |
|---|---|---|---|
house-oversight |
House | typed | yes |
hsgac |
Senate | typed | yes |
senate-commerce |
Senate | typed | yes |
senate-judiciary |
Senate | flat | no |
senate-foreign-relations |
Senate | flat | no |
senate-help |
Senate | flat | no |
senate-aging |
Senate | flat | no |
~40,600 documents indexed. All 30 major House and Senate committees were surveyed on 2026-07-30: most House committees publish no sitemap at all, so they cannot be added without a new adapter.
flat committees publish a sitemap with no dates. Their results carry last_updated: null, sort alphabetically rather than by recency, and refuse since instead of applying it to undated rows — which would return pre-cutoff documents as though they had passed the filter. The response says so via dates_unavailable.
Not every committee publishes every type. A doc_type a committee does not publish returns an explicit unavailable_types and a note, never an empty list that would read as "nothing on the subject".
Coverage and limits
- Search matches the document headline/slug, not body or PDF text. Searching 40,000 pages of body text needs a hosted index. A null result means no headline matched — not that the subject is absent.
- PDF text is not extracted. Attachments come back as url + filename + bytes. Quote the page text and cite the attachment as a link; do not represent a PDF as having been read.
- Some pages are pure file wrappers with no prose of their own; those return empty text plus a
no_prose_noterather than the site footer. get_committee_documentfetches a caller-supplied URL, so it is host-allowlisted to the seven committee domains.
A deliberate omission
No person→mentions tool, for the reasons set out in congressional-documents: a name is not an identifier, and being named in a committee document is not an allegation.
Data sources
Each committee's Yoast sitemap index or sitemap.xml, e.g.:
https://oversight.house.gov/sitemap_index.xmlhttps://www.hsgac.senate.gov/sitemap_index.xmlhttps://www.judiciary.senate.gov/sitemap.xml
All are served under a robots.txt with an empty Disallow: — crawling explicitly permitted.
Quick Start
Add to your MCP client (Claude Desktop, Cursor, Windsurf, etc.):
{
"mcpServers": {
"committee-releases": {
"url": "https://gateway.pipeworx.io/committee-releases/mcp"
}
}
}
Or connect to the full Pipeworx gateway for access to all 1394+ data sources:
{
"mcpServers": {
"pipeworx": {
"url": "https://gateway.pipeworx.io/mcp"
}
}
}
Using with ask_pipeworx
Instead of calling tools directly, you can ask questions in plain English:
ask_pipeworx({ question: "your question about Committee Releases data" })
The gateway picks the right tool and fills the arguments automatically.
More
License
MIT
Recommended Servers
playwright-mcp
A Model Context Protocol server that enables LLMs to interact with web pages through structured accessibility snapshots without requiring vision models or screenshots.
Magic Component Platform (MCP)
An AI-powered tool that generates modern UI components from natural language descriptions, integrating with popular IDEs to streamline UI development workflow.
Audiense Insights MCP Server
Enables interaction with Audiense Insights accounts via the Model Context Protocol, facilitating the extraction and analysis of marketing insights and audience data including demographics, behavior, and influencer engagement.
VeyraX MCP
Single MCP tool to connect all your favorite tools: Gmail, Calendar and 40 more.
graphlit-mcp-server
The Model Context Protocol (MCP) Server enables integration between MCP clients and the Graphlit service. Ingest anything from Slack to Gmail to podcast feeds, in addition to web crawling, into a Graphlit project - and then retrieve relevant contents from the MCP client.
Kagi MCP Server
An MCP server that integrates Kagi search capabilities with Claude AI, enabling Claude to perform real-time web searches when answering questions that require up-to-date information.
E2B
Using MCP to run code via e2b.
Neon Database
MCP server for interacting with Neon Management API and databases
Exa Search
A Model Context Protocol (MCP) server lets AI assistants like Claude use the Exa AI Search API for web searches. This setup allows AI models to get real-time web information in a safe and controlled way.
Qdrant Server
This repository is an example of how to create a MCP server for Qdrant, a vector search engine.