House Clerk roll-call votes

planned Legislative · Tier 1 · XML index · Clerk of the House of Representatives

What this source is

Every House roll-call vote as XML at clerk.house.gov/evs/{year}/roll{NNN}.xml with a yearly index — published same-day, examples back to 2001. Pairs with the Congress.gov API's beta house-vote endpoint as a cross-check.

Identity and registry record

Registry id
house-clerk-votes
Agency / parent organization
Clerk of the House of Representatives
Branch
legislative
Type
XML index
Status
planned
Tier
1
URL (home)
https://clerk.house.gov/Votes (opens in a new tab)
URL (index)
https://clerk.house.gov/evs/2026/index.asp (opens in a new tab)
Registered
2026-07-28
Registry notes
Vote XML files verifiably exist and are indexed (2026-07-28). The formal schema documentation lives on xml.house.gov, which 403'd research fetchers — read it via our identified client at probe. 2026-07-31, live: the per-vote XML is confirmed (evs/2026/roll283.xml → 200 text/xml, 82,500 bytes, full metadata + member positions) and clerk.house.gov publishes NO robots.txt (404, nothing disallowed), but the YEAR INDEX IS HTML, NOT XML — evs/2026/index.asp is a 7 KB <TABLE> of ~15 recent votes linking to cgi-bin/vote.asp, evs/2026/index.xml is 404, and clerk.house.gov/Votes is a 249 KB JavaScript application. So this entry needs the html-index adapter (plan Phase 5), not the xml-index one that activated senate-xml; it stays planned rather than ship half-understood. Type stays xml-index because the per-vote records are XML; revisit the type when the index adapter lands.

How we ingest it

Channel
XML index
Method
Would walk the numeric roll-call sequence from a watermark, fetching new vote XML files.
Request budget
the agency class: at most 1,500 requests per day shared across every agency web source, counted from the fetch log (failed requests count too)
Politeness
robots.txt is honored as observed — including each host's crawl-delay, exactly — and every request identifies itself as fapd/0.1 (Free Agentic Publication Digester; +https://fapd.info/bot.html; contact: hustleyourcity@gmail.com); a refusal is recorded, never evaded
Capture and hash
captured raw content is hashed (SHA-256) into the day's committed provenance manifest, hash-chained day to day (PROVENANCE.md)

Ingestion health

Not ingested: the registry status of this source is planned.

Ingestion statistics

These figures describe this project's ingestion of this source — items we recorded and requests we made — and nothing else. They are not a measurement of the publisher.

Not ingested: the registry status of this source is planned. Ingestion statistics are measured for active sources only.

All time

No requests to clerk.house.gov are recorded in the request log.

Last 30 days, day by day

No requests and no items were recorded in the last 30 days, so there is nothing to chart.