Capture what went wrong
Execution traces give the teacher context about failed steps and recurring patterns.
Manifest turns lessons from failed coding-agent runs into reusable, tested routines for small local models.
Explore the prototype ↗A research workflow, implemented in the public repository.
Small models can repeat the same process mistakes across coding tasks. Manifest explores whether a stronger teacher can turn those failures into explicit control-flow routines that a local agent can reuse.
Execution traces give the teacher context about failed steps and recurring patterns.
Proposed routines pass schema and code validation, repair, and regression checks before joining the harness.
The Warden uses a permission manifest and a kernel sandbox to constrain generated routines. Accepted routines can run with a local model through Ollama.
The public project includes a CLI, a GUI, and an evaluation pipeline. The current teacher uses DeepSeek. We plan to integrate Claude for trace analysis and routine generation, then compare task success, regressions, inference cost, and sandbox compliance using fixed task splits and budgets.
Manifest is an early-stage research prototype. Claude integration is planned; production readiness and performance gains have not been established.
Manifest is a bootstrapped venture based in India, founded in October 2026. We are exploring developer tooling that makes small local coding agents more dependable through tested, reusable routines.