33
Rust harnesses and open control planes that scale agent fleets
Power users are wiring clone-and-run sandboxes, memory-efficient loops, and verification gates into custom fleets. These repos prioritize systems-language runtimes, declarative orchestration, and independent judges over closed IDE chatbots.
01 — Power tools
The stack
- 01harnessGitHubTrueForge↗
TrueFoundry
Open-source agent harness runtime that handles model calls, MCP tools, skills, sandboxing, approvals, context compaction, and session state with chat UI plus TypeScript SDK.
Use caseUse it to deploy cost-controlled multi-step agents locally or hosted while swapping any model provider and MCP servers.
harnessmcpsandboxsdk - 02toolGitHubjcode↗
1jehuang
High-performance Rust coding agent harness optimized for low RAM, fast boot, semantic memory graphs, side panels, and native swarm collaboration across sessions.
Use caseUse it to run dozens of parallel coding sessions on limited hardware with automatic conflict notification and memory recall.
rustharnessswarmmemory - 03toolGitHubCodewhale↗
Hmbown
Open-source Rust terminal coding agent with multi-model support, fleets, MCP, hooks, approval modes, snapshots, and local-first extensibility.
Use caseUse it to coordinate agent teams on repo tasks with explicit plan or full-access modes and undoable workspace restores.
rustclifleetmcp - 04harnessGitHubPi↗
earendil-works
Minimal extensible coding agent harness with unified multi-provider LLM API, agent core loop, TUI, skills, extensions, and session persistence.
Use caseUse it to build or extend a lean coding CLI that stays small at core while adding TypeScript packages for custom tools and themes.
harnesscliextensiblemulti-provider - 05harnessGitHubSWE-AF↗
Agent-Field
Autonomous software engineering fleet runtime that spins up planners, coders, reviewers, and testers to ship production-grade PRs from a single goal call.
Use caseUse it to orchestrate multi-repo builds with dependency scheduling, git worktrees, and adaptive factory control for end-to-end delivery.
fleetorchestrationswemulti-agent - 06repoGitHubAgentic Harness Engineering (AHE)↗
china-qijizhifeng
Observability-driven system that automatically evolves coding-agent harness components such as prompts, tools, middleware, and memory while holding the base model fixed.
Use caseUse it to iterate harness configs over evaluate-analyze-improve loops and freeze transferable components that lift Terminal-Bench scores.
harness-engineeringobservabilityevolutioneval - 07harnessGitHubagentic-sandbox↗
jmagly
Self-hostable runtime for persistent autonomous coding agents using KVM-isolated VMs or rootless containers with A2A protocol, dashboard, and virtiofs storage.
Use caseUse it to run long-lived agent sessions on your own hardware with signed AgentCard discovery and mission dispatch without a hosted control plane.
sandboxkvmself-hostedpersistent - 08harnessGitHubAgentTier↗
agenttier
Kubernetes-native sandbox platform providing isolated persistent environments for AI agents and developers via CRDs with gVisor options and interactive access.
Use caseUse it to provision secure sandboxes for coding agents or teams with declarative lifecycle, network isolation, and multi-agent orchestration.
kubernetessandboxcrdisolation - 09harnessGitHubagent-harness↗
adambossy
Provider-agnostic modular Python agent harness with Protocols for models, tools, sessions, and sandboxes plus async typed core under 3k LOC.
Use caseUse it to assemble a custom loop that plugs Anthropic, OpenAI, Modal sandboxes, or Redis sessions without god-classes.
pythonmodularprovider-agnosticsandbox - 10harnessGitHubSWE Forge↗
joacod
Portable opt-in workflow orchestration that sits above coding harnesses to turn tickets into bounded inspect-plan-implement-verify-review deliveries.
Use caseUse it to add evidence-backed SOLO or SUBAGENTS topologies over Pi, OpenCode, or experimental Claude Code adapters without replacing the base harness.
workfloworchestrationverifyharness-agnostic - 11toolGitHubopenteams↗
openteams-lab
Local-first AI desktop app that plans, builds, and ships with a controllable team of coding agents including many popular CLIs plus a bundled CLI.
Use caseUse it to assign roles across Claude Code, Codex, Pi, and others inside one workspace with budget caps and reviewer loops.
desktopmulti-agentlocal-firstorchestration - 12toolGitHubCodey↗
its-ahoh
Multi-agent workbench that orchestrates Claude Code, Codex, OpenCode, and Pi from a native macOS app, chat platforms, or voice with per-project workspaces.
Use caseUse it to run parallel agents on the same prompt for comparison or define worker teams with flow graphs and messaging channels.
workbenchmulti-agentmacosparallel - 13repoGitHubSWE-agent↗
SWE-agent
LM-driven harness built for SWE-bench that provides edit state, command execution, and an issue-focused loop as a reference agent-computer interface.
Use caseUse it to clone a battle-tested issue-to-PR agent stack and adapt the YAML config for custom benchmarks or cybersecurity tasks.
swe-benchharnessresearch-originpython - 14repoGitHubremote-swe-agents↗
aws-samples
Self-hosted fully open-source autonomous SWE agent example on AWS that works in dedicated cloud environments and opens PRs.
Use caseUse it to run laptop-free coding agents that clone, edit, test, and submit pull requests inside isolated AWS workspaces.
awsautonomousswecloud - 15harnessGitHubAgentic Harness↗
moortekweb-art
Self-hosted completion gate that wraps coding agents with independent verification, durable evidence, CLI, and local GUI before reporting done.
Use caseUse it to force an external check and report.md before any agent can claim a goal is finished across compatible CLIs.
verificationgatecompletionself-hosted - 16harnessGitHubHarness↗
majiayu000
Rust control plane for fleets of parallel coding agents with orchestration, policy enforcement, cross-agent review, and observability.
Use caseUse it to assign work, police permissions, and run independent reviews across Claude Code or Codex adapters in long-running fleets.
rustcontrol-planefleetpolicy - 17harnessGitHubOpenHarness↗
maisieyang
Local-first Python control plane for coding agents featuring Default, Plan, externally judged Goals, sandboxed execution, durable context, plugins, and evals.
Use caseUse it to drive REPL sessions with independent judges that decide continue, complete, or pause based on artifacts rather than agent claims.
pythoncontrol-planejudgesandbox - 18harnessGitHubHAR↗
os-factory
Open agent harness CLI plus MCP that runs coding agents in isolated worktrees with deterministic launch, verify, and teardown stages.
Use caseUse it to turn any repo into verified software-factory slots that wrap Claude Code, Cursor, or Codex with project-native checks.
worktreeverifymcpcli - 19repoGitHubBlaxel Sandbox↗
blaxel-ai
Open-source microVM sandbox templates and hub for persistent agent compute with standby resume, MCP tools, filesystem, and process APIs.
Use caseUse it to give agents dedicated Linux environments that resume in tens of milliseconds while keeping full state across idle periods.
sandboxmicrovmmcppersistent - 20repoGitHubSandboxer↗
hyperterse
Unified client libraries in Go, Python, and TypeScript that talk directly to E2B, Daytona, Blaxel, Runloop, Fly, or local Docker sandboxes.
Use caseUse it to switch sandbox providers behind one mental model for create, exec, filesystem, and teardown without rewriting agent code.
sandboxmulti-providersdkabstraction
02 — In the wild
Articles and releases
- releaseTrueFoundry Blog / VentureBeatTrueForge open-sources its production agent harness with 30 percent lower cost than Claude Managed Agents↗
TrueFoundry
TrueFoundry released TrueForge under MIT as a vendor-neutral harness with MCP, skills, sandboxing, and compaction. Benchmarks show matching quality at substantially lower token and dollar cost versus managed alternatives on identical tasks.
releaseharnessopen-sourcecostShipped
- articleHacker NewsShow HN: Supafork lets you store, share, and fork agent sessions across harnesses↗
supafork
Supafork provides a GitHub-like layer for AI agent sessions so users can save complete runs from Claude Code, Codex, Pi, and others, organize collections, and import across harnesses. The launch post highlights CLI import plus public or private forking of workflows.
articlesessionsinteroperabilityshow-hnRead
- articleHacker News / mouse.devMouse tops FrontierHarness leaderboard using completion loops on Kimi K3↗
Aeroi
Mouse, built on a customized OpenCode base with specialized rules, skills, and loops, scored highest pass rate and first on time among twelve harnesses all running the same Kimi K3 model. The writeup details overnight Night Shift mode and the completion-loop changes that drove the result.
articlebenchmarkharnessleaderboardRead