
Pairing each generated claim with an evidence packet and routing unsupported ones back nearly doubles relation accuracy over non- baselines (0.676 vs 0.383) on blind labels from AVeriTeC, CLIMATE-FEVER and SciFact.
articleECLoop: Evidence-Conditioned Execution Layer for Coding AgentsYisen Xu, Chenglin Li, Zehao Wang, Jinqiu Yang, Tse-Hsun Chen
articleEA-Graph: Artifact-Anchored Verification Memory for Coding Agents under Upstream DriftHwai-Jung Hsu, Cheng-Jan Chi, Hanna Everett
articleSARC-DQ: Runtime Data-Quality Gating for Agentic AI: Silent Evidence Defects, the Incompetence Shield, and Downstream-Only RemediationGaston BesansonChecking sign-in…
Loading comments…