Can you trust it when AI says "it's done"?

Built Foreman, which uses Jev to continuously check whether a coding agent has drifted, stalled, or finished.
The original post, translated
Can you trust it when AI says "it's done"? This project gives the Coding Agent a supervisor. The project is called Foreman. Codex / OpenCode is responsible for writing code, while another model, Jev, continuously checks nearby: whether it has gone off track, whether it is stuck, whether there are enough tests, and whether the requirements are actually finished, and then the program decides the next step according to the rules. Core highlights: 1️⃣
Show the original post in its source language ▾
AI 说“写好了”就能信?这个项目给 Coding Agent 配了个监工。 项目叫 Foreman。Codex / OpenCode 负责写代码,另一个模型 Jev 在旁边持续检查:有没有跑偏、是不是卡住、测试够不够、需求到底做完没有,再由程序按规则决定下一步。 核心亮点: 1️⃣
The post above is a machine translation from zh; the untranslated text is in the fold-out.
Engagement when collected
Numbers are a snapshot taken from X when the case was added to the library (schema v1, collected 2026-09-19); they will not match today.
Where this case fits
Filed under agents & workflows, coding & developer tools. In the pattern Jev is built for, the model answers a bounded question per step — and ordinary code acts on the answer, because the answer is already a value rather than a paragraph. Other posts in the same family are on the agents & workflows page.
Related Jev cases
found the perfect use case for @typesafeai Jev: instant
Agents
Breaking: Browser Use + Jev = Ultrafast Findings flights
AgentsBrowser agents
The gains aren’t free: Jev can't generate text Comparing
Dev toolsClassification
Jev case: in 40 seconds it broke down 724 live ads from 37
WebsitesAgents
Keep browsing: all 1173 Jev cases · more from @TiedGST · builders · what Jev is
Last updated: 2026-09-22 · sources & corrections · every card links to its author's original post