ProvenanceGuard checks LLM agent answers for correct sourcing
Hugging Face researchers have introduced ProvenanceGuard, a new system that checks whether large language model (LLM) agents are attributing facts to the right sources, not just whether the facts are correct. Unlike existing tools that mix all evidence together, ProvenanceGuard keeps track of where each piece of information came from and flags answers that cite the wrong source—even if the information itself is accurate somewhere in the evidence.
- ProvenanceGuard works after the LLM produces an answer
- Checks if each claim cites the correct source tool
- Prevents cross-source conflation errors
- Maintains source identity throughout verification
- No need to retrain the underlying LLM agent
Sources covering this
More in AI
OpenAI launches Dots AI agents amid delayed model update
OpenAI has released Dots, its new always-on AI assistant designed to autonomously handle tasks for users.
OpenAI rolls out Dots, always-on AI agents for work tasks
OpenAI has launched Dots, always-available AI assistants designed to handle projects, tasks, and reminders across apps like Slack,…
Nvidia releases Open Agent Safety Platform to contain rogue AI
Nvidia has launched the Open Agent Safety Platform, aimed at containing AI agents that attempt to circumvent their boundaries.
OpenAI cancels GPT-6.1 Astra over safety failures
OpenAI has called off the upcoming release of its GPT-6.1 Astra model after internal tests flagged safety problems.