Check out the paper for all the details: arxiv.org/abs/2511.18749 Thanks to my collaborators @yang3kc.bsky.social @harryyan.bsky.social and @fil.bsky.social .
Large Language Models Require Curated Context for Reliable Political Fact-Checking -- Even with Reasoning and Web Search
Large language models (LLMs) have raised hopes for automated end-to-end fact-checking, but prior studies report mixed results. As mainstream chatbots increasingly ship with reasoning capabilities and ...
arxiv.org