Skip to content

Entry

Effective Harnesses for Long-Running Agents

Appears in 4 awesome lists

Anthropic's pattern for maintaining agent progress across multiple context windows: an initializer agent sets up the environment once and hands off to a coding agent that makes incremental progress each session. The structured handoff mechanism — feature lists, git commits, and test gates as…

Open anthropic.com

Found in these lists

Awesome Artificial Intelligence

Section: Durable and asynchronous agents · Patterns for agents that make progress across multiple context windows and recover from failure.

FreshScore 86

Awesome Harness Engineering

Section: Planning & Task Decomposition · Anthropic's pattern for maintaining agent progress across multiple context windows: an initializer agent sets up the environment once and hands off to a coding agent that makes incremental progress each session. The structured handoff mechanism — feature lists, git commits, and test gates as…

FreshScore 88

Awesome Prompts

Section: Harness Engineering · Long-running agent design

FreshScore 90

Best Of Agent Harnesses

Section: Related Resources · Session bridging, feature lists, incremental progress, testing

FreshScore 84

ECC

Affaan Momin's agent-harness operating system: 68 specialized agents, 286 skills, hooks, memory, continuous learning, and AgentShield security scanning across Claude Code, Codex, Cursor, OpenCode, and other harnesses. The clearest open-source example of packaging an end-to-end engineering workflow…

In 7 listsDetails

langchain-ai/deepagents

LangChain's batteries-included agent harness (released April 2026) with built-in planning, filesystem tools, shell access, sub-agents, and auto-summarization. The clearest open-source demonstration of how a general-purpose coding agent harness can be made ready-to-run out of the box while…

In 7 listsDetails

Harness Engineering

OpenAI's framing of harness engineering as a discipline: how to design the scaffolding that lets Codex and similar agents operate reliably in an agent-first world.

In 4 listsDetails

Mirage

Swaps the filesystem and bash providers for a mirage virtual workspace: file tools and shell commands run over mounted resources (RAM, S3, Redis, Slack, Gmail, Notion, Postgres) instead of the host disk, with per-mount read/write/exec modes, per-command sandbox routing (monty, pyodide, quickjs in…

In 3 lists

Loop Engineering

"Stop prompting. Design the loop." — practical patterns, starters & CLI (loop-audit, loop-init, loop-cost) for systems that discover work, hand it to agents, verify results, and persist state across Claude Code, Codex, Grok, and OpenCode; report-only week one, scores loops on a "Loop Ready" rubric…

In 3 lists

Spec Kit

GitHub's open-source toolkit for spec-driven development: spec.md → plan.md → tasks.md artifacts convert natural-language intent into a machine-checkable plan before any code is written, with /specify, /plan, /tasks, and /implement commands that run across Claude Code, Copilot, Codex, and Gemini…

In 3 lists

TaskWeaver

Code-first task decomposition framework with a planner/executor split and a plugin system for injecting domain knowledge into the planning layer. The most complete reference implementation of plan-then-execute with stateful task tracking.

In 3 lists

Building a C Compiler with a Team of Parallel Claudes

Anthropic's account of coordinating 16 Claude instances in parallel on a shared git repo without a central orchestrator: agents claim tasks via files in current_tasks/, git forces collision resolution naturally, and a continuous restart loop spawns fresh sessions that resume where predecessors…

In 2 lists