Do AI chatbots give voters accurate election information?
Researchers tested whether ChatGPT and Google AI could reliably answer common voter questions. The stakes matter because voters increasingly turn to AI for guidance on where and how to vote.
States United Democracy Center ran two rounds of empirical testing of ChatGPT and Google AI on common voter questions, in late 2025 and early 2026, ahead of the 2026 midterms. In the "preliminary" round, error rates were 6.9% (Google AI) and 8.2% (ChatGPT); by the "primary" round, "the rate of verifiable factual errors fell to 0% across both platforms." The study's bottom line, though, is that this accuracy gain did not make the platforms adequate substitutes for official sources: "voters seeking election information on leading AI platforms are inconsistently directed to the authoritative information on their state's official election website." ChatGPT mentioned a state election site in 39.4% of responses and Google AI in 55.6% — "less than 50% of the time" overall, and "for some questions, they almost never did."
The report is explicit that zero measured errors is a narrower claim than it sounds: "It is worth being precise about what a 0% error rate does and does not establish... It does not mean, however, that the responses were complete." ChatGPT returned incomplete candidate lists for 88.9% of gubernatorial queries, which the study attributes to a structural mismatch — "AI systems trained on historical data and reliant on periodic retrieval were structurally ill-suited to track real-time candidate filings" like an actively changing field such as Arizona's. Compounding this, candidate questions relied heavily on Wikipedia (12.3% of all 3,481 cited links, and the dominant source for candidate questions specifically), which the study flags as "openly editable and not authoritative for time-sensitive or jurisdiction-specific election information." Google AI also changed its output format mid-study, on or around Feb. 2, 2026, replacing prose summaries with links-only answers; asked why, Google attributed the change both to reducing hallucinations and to "making AI search more monetizable" — after which Google AI directed voters to a state site in none of its responses.
This sits alongside Did ChatGPT cause Stack Overflow posting to decline? as a parallel case of a chatbot interposing itself between a user and the authoritative source rather than routing to it — there, a drop in traffic to the platform that held the answers; here, a measured shortfall in links back to the office that administers the information. It also echoes How much GPT-written scholarship reaches Google Scholar undetected? in flagging degraded provenance in a nominally authoritative channel, though the two findings are not the same mechanism: Haider's concern is undisclosed AI content entering a scholarly index, while States United's is a chatbot citing an editable encyclopedia and unvetted platforms (Reddit, YouTube) for facts that only an election office can certify.
The study covers two AI platforms, a defined question set, and three states over several weeks in round two, so it does not establish how other assistants, other states, or other election cycles would score, nor does it isolate how much of the completeness gap is retrieval architecture versus deliberate design choice (the "monetizable" framing for Google's format change is Google's own stated explanation, not something the study verifies independently). Within those bounds, the finding supports treating chatbot accuracy and chatbot completeness as separate, non-substitutable measures of voter utility — a system can reach zero measured factual errors while still failing the single measure the study treats as most consequential, directing the user onward to the source that is actually accountable for the answer.
Inquiring lines that read this note 10
This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.
How can AI systems reliably guide voters without introducing political bias?- Does ChatGPT displace search engines or question-and-answer platforms?
- Do chatbots actually give consistent voting recommendations regardless of user input?
- What political information quality issues arise when AI tools guide voting?
- What accuracy do AI chatbots actually provide on election topics?
- Should election regulators require chatbots to refuse voting advice entirely?
- Do other AI assistants perform similarly on voter election questions?
- Can AI systems stay current with real-time candidate filings across states?
- How many Dutch voters actually plan to use AI for voting advice?
- What makes voting-advice tools like Kieskompas more reliable than chatbots?
- Can voters distinguish between confident chatbot answers and accurate ones?
Related concepts in this collection 4
This note in its neighbourhood — explore the map, then jump to a related concept in the list below.
Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph
-
Did ChatGPT cause Stack Overflow posting to decline?
Researchers used a difference-in-differences model to test whether public programming Q&A posting fell after ChatGPT's release. This matters because it could signal whether AI tools are shifting knowledge from public commons to private use.
both show chatbots standing between users and the authoritative source rather than routing to it
-
How much GPT-written scholarship reaches Google Scholar undetected?
Haider et al. searched Google Scholar for telltale ChatGPT phrases to estimate how many papers contain undisclosed AI authorship, especially in policy-relevant fields. Understanding prevalence matters because lay readers—politicians, patients, students—may treat these papers as credible research.
parallel concern about degraded provenance inside a channel voters or researchers treat as authoritative
-
Do AI assistants reliably answer questions about news?
A major cross-country study tested whether AI assistants like ChatGPT and Gemini accurately handle news queries. Understanding this matters because many people may turn to these tools for current events.
Evidence for A: a cross-market study of 3,000+ AI news answers finds sourcing failures recur broadly, matching A's sourcing gap
-
Will voters actually use AI chatbots for election information?
Explores how many voters plan to rely on AI chatbots versus traditional news sources for 2026 election information, and whether stated intent reflects actual behavior or real-world impact.
Qualifies A: finds only 15% of voters plan to use AI chatbots for election info, limiting practical impact of accuracy
Related papers in this collection 8
Papers most semantically related to this note, ranked by cosine similarity in the embedding space.
- AI and Elections: How Well Do AI Platforms Answer Voter Questions?
- Who's Asking AI About the 2026 Election?
- Artificial intelligence is ineffective and potentially harmful for fact checking
- Auditing Political Alignment in LLM Assistants: Engagement, Stance, and User Identity
- Dutch privacy watchdog warns against using AI chatbots for voting advice
- The Decision to Verify: How Warmth and User Characteristics Shape Reliance on Conversational Agents for Information Search
- How AI Is Changing Search Behaviors
- Digital News Report 2026
Original note title
States United finds ChatGPT and Google AI hit 0% factual error on voter questions but rarely direct voters to official state sites