Back to collection
Research

VEX 0.2 Chronos Targets More Reliable Agent Execution

Long-task performance needs completion evidence, not fluent output.

Two people stand beside a workshop bench and look at a red emergency stop button on a white box.
Technical illustration

VEX 0.2 Chronos targets faster execution, steadier long tasks and fewer wasted steps. The meaningful distinction is between an agent's conversational output and completed work: repeated invalid actions or silent failures can consume capacity without producing a usable result. The accompanying report of knowledge queries and reasoning paid in TRAC highlights a second boundary, from deciding to spending. Completion, intervention frequency, payment accuracy and recoverability are more useful evaluation criteria than eloquence or transaction volume alone.

References