Apodex 1.1 review: Statement Review independently verifies agent conclusions

aakashgupta · x · 2026-09-09

The author argues Apodex 1.1's sleeper feature is Statement Review: generation and review are separated, key conclusions get independently checked before delivery, and thin evidence, citation mismatches, or conflicting numbers are flagged with corrections and an audit trail — crucial when agent output feeds real decisions.

The first thing to test is mid-task intervention: drop a new file or changed requirement into a running job and only affected parts get replanned while valid work is kept (demo: a global EV battery supplier risk assessment that absorbs new documents mid-flight). Under the hood is an asynchronous Agent Team in Deep Discover mode — the model decides how to split tasks, how many subagents to spawn, and when to consolidate, with subagents streaming results into shared task state on a live task board. Complete capability is defined as six things: understand the goal, act in a real environment, hold state over long runs, absorb mid-execution feedback, repair and continue, deliver verifiable results — "move reasoning out of the report and into the execution of real tasks." The stack is open source (FrontierAgent CLI framework, one command on macOS/Linux, no Docker, fully local with mini weights).

Related event: Apodex 1.1 Ships: Agent System Measured by End-to-End Deliverables(10 posts)→

Original post →

More from coding & agent

coding & agent channel →