Baseline evidence pack
Run one fixed task before adding scaffolding and separate model, context, state, control, and feedback failures.
Output: Baseline receipt, task contract, and first-intervention decision.
For software engineers, technical leads, and developer-platform owners
Turn a coding agent into a repository system that can understand, verify, resume, and hand off real work.
Eleven field modules retrofit one brownfield application with task contracts, repository knowledge, isolated environments, durable state, agent-readable feedback, mechanical controls, verification, recovery, and measured scaling decisions.
Open module 00Foundation → lab → project
Each lesson produces an artifact, tests it against observable criteria, then carries it into the next decision. Enter where your workflow is stuck or follow the complete sequence.
Hands-on projects
Run one fixed task before adding scaffolding and separate model, context, state, control, and feedback failures.
Output: Baseline receipt, task contract, and first-intervention decision.
Build the knowledge map, deterministic initializer, isolated worktree resources, and fresh-session handoff.
Output: Repository map, startup receipt, feature ledger, and clean handoff.
Expose runtime feedback, enforce architecture and permissions, grade the real user path, and recover bounded failures.
Output: Mechanical rules, eval suite, trace, budgets, and recovery runbook.
Use the full harness on a holdout feature and remove any layer that fails to produce measured value.
Output: Repo harness kit, Harness Card, before-and-after report, and go or no-go decision.
Tools, field guides, and services
The runtime, control, observability, and recovery layer around an agent loop.
Open resourceField guidePlans and acceptance criteria as durable, reviewable engineering artifacts.
Open resourceField guideWorktrees, ephemeral environments, and blast-radius control for long-running agents.
Open resourcePlaybookA maintained source radar for foundations, evals, safe autonomy, and reference harnesses.
Open resourceWhen a local harness meets a real codebase
Tenten can review repository legibility, permissions, evaluator coverage, worktree isolation, recovery, and rollout evidence before your team increases agent autonomy.