#Verification
5 articles with this tag

Dotta on Defining 'Done' for AI Agents and Paperclip's Liveness Model
Dotta, creator of Paperclip, explains how to define "done" for AI agents, emphasizing a "reliance claim" model over a simple boolean, balancing liveness and assurance.

AI Agents: The Rebuilt CI/CD Pitfalls
Sumaiya Shrabony warns that solo AI agent builders often recreate flawed CI/CD processes, leading to issues like voice drift and missing verification.

LLM Verification: A New Scaling Axis
LLM-as-a-Verifier redefines LLM scaling by treating verification as a new axis, offering continuous scores for enhanced accuracy and efficiency across agentic tasks.

Scaling AI Beyond Informal: Axiom Math's Carina Hong
Carina Hong of Axiom Math discusses scaling AI through formal verification, aiming to build reliable and collaborative systems.

Causal Verification for Reliable Tool Use
CIVeX, a causal intervention verifier, ensures reliable tool use by focusing on intervention identifiability, not just action validity, achieving zero false executions in adversarial settings.

