~/reading

33

Issue 33/September 6, 2026/20 tools · 3 headlines

Rust harnesses and open control planes that scale agent fleets

Power users are wiring clone-and-run sandboxes, memory-efficient loops, and verification gates into custom fleets. These repos prioritize systems-language runtimes, declarative orchestration, and independent judges over closed IDE chatbots.

01 — Power tools

The stack

  1. 01harnessGitHub
    TrueForge

    TrueFoundry

    Open-source agent harness runtime that handles model calls, MCP tools, skills, sandboxing, approvals, context compaction, and session state with chat UI plus TypeScript SDK.

    Use caseUse it to deploy cost-controlled multi-step agents locally or hosted while swapping any model provider and MCP servers.

    harnessmcpsandboxsdk
  2. 02toolGitHub
    jcode

    1jehuang

    High-performance Rust coding agent harness optimized for low RAM, fast boot, semantic memory graphs, side panels, and native swarm collaboration across sessions.

    Use caseUse it to run dozens of parallel coding sessions on limited hardware with automatic conflict notification and memory recall.

    rustharnessswarmmemory
  3. 03toolGitHub
    Codewhale

    Hmbown

    Open-source Rust terminal coding agent with multi-model support, fleets, MCP, hooks, approval modes, snapshots, and local-first extensibility.

    Use caseUse it to coordinate agent teams on repo tasks with explicit plan or full-access modes and undoable workspace restores.

    rustclifleetmcp
  4. 04harnessGitHub
    Pi

    earendil-works

    Minimal extensible coding agent harness with unified multi-provider LLM API, agent core loop, TUI, skills, extensions, and session persistence.

    Use caseUse it to build or extend a lean coding CLI that stays small at core while adding TypeScript packages for custom tools and themes.

    harnesscliextensiblemulti-provider
  5. 05harnessGitHub
    SWE-AF

    Agent-Field

    Autonomous software engineering fleet runtime that spins up planners, coders, reviewers, and testers to ship production-grade PRs from a single goal call.

    Use caseUse it to orchestrate multi-repo builds with dependency scheduling, git worktrees, and adaptive factory control for end-to-end delivery.

    fleetorchestrationswemulti-agent
  6. 06repoGitHub
    Agentic Harness Engineering (AHE)

    china-qijizhifeng

    Observability-driven system that automatically evolves coding-agent harness components such as prompts, tools, middleware, and memory while holding the base model fixed.

    Use caseUse it to iterate harness configs over evaluate-analyze-improve loops and freeze transferable components that lift Terminal-Bench scores.

    harness-engineeringobservabilityevolutioneval
  7. 07harnessGitHub
    agentic-sandbox

    jmagly

    Self-hostable runtime for persistent autonomous coding agents using KVM-isolated VMs or rootless containers with A2A protocol, dashboard, and virtiofs storage.

    Use caseUse it to run long-lived agent sessions on your own hardware with signed AgentCard discovery and mission dispatch without a hosted control plane.

    sandboxkvmself-hostedpersistent
  8. 08harnessGitHub
    AgentTier

    agenttier

    Kubernetes-native sandbox platform providing isolated persistent environments for AI agents and developers via CRDs with gVisor options and interactive access.

    Use caseUse it to provision secure sandboxes for coding agents or teams with declarative lifecycle, network isolation, and multi-agent orchestration.

    kubernetessandboxcrdisolation
  9. 09harnessGitHub
    agent-harness

    adambossy

    Provider-agnostic modular Python agent harness with Protocols for models, tools, sessions, and sandboxes plus async typed core under 3k LOC.

    Use caseUse it to assemble a custom loop that plugs Anthropic, OpenAI, Modal sandboxes, or Redis sessions without god-classes.

    pythonmodularprovider-agnosticsandbox
  10. 10harnessGitHub
    SWE Forge

    joacod

    Portable opt-in workflow orchestration that sits above coding harnesses to turn tickets into bounded inspect-plan-implement-verify-review deliveries.

    Use caseUse it to add evidence-backed SOLO or SUBAGENTS topologies over Pi, OpenCode, or experimental Claude Code adapters without replacing the base harness.

    workfloworchestrationverifyharness-agnostic
  11. 11toolGitHub
    openteams

    openteams-lab

    Local-first AI desktop app that plans, builds, and ships with a controllable team of coding agents including many popular CLIs plus a bundled CLI.

    Use caseUse it to assign roles across Claude Code, Codex, Pi, and others inside one workspace with budget caps and reviewer loops.

    desktopmulti-agentlocal-firstorchestration
  12. 12toolGitHub
    Codey

    its-ahoh

    Multi-agent workbench that orchestrates Claude Code, Codex, OpenCode, and Pi from a native macOS app, chat platforms, or voice with per-project workspaces.

    Use caseUse it to run parallel agents on the same prompt for comparison or define worker teams with flow graphs and messaging channels.

    workbenchmulti-agentmacosparallel
  13. 13repoGitHub
    SWE-agent

    SWE-agent

    LM-driven harness built for SWE-bench that provides edit state, command execution, and an issue-focused loop as a reference agent-computer interface.

    Use caseUse it to clone a battle-tested issue-to-PR agent stack and adapt the YAML config for custom benchmarks or cybersecurity tasks.

    swe-benchharnessresearch-originpython
  14. 14repoGitHub
    remote-swe-agents

    aws-samples

    Self-hosted fully open-source autonomous SWE agent example on AWS that works in dedicated cloud environments and opens PRs.

    Use caseUse it to run laptop-free coding agents that clone, edit, test, and submit pull requests inside isolated AWS workspaces.

    awsautonomousswecloud
  15. 15harnessGitHub
    Agentic Harness

    moortekweb-art

    Self-hosted completion gate that wraps coding agents with independent verification, durable evidence, CLI, and local GUI before reporting done.

    Use caseUse it to force an external check and report.md before any agent can claim a goal is finished across compatible CLIs.

    verificationgatecompletionself-hosted
  16. 16harnessGitHub
    Harness

    majiayu000

    Rust control plane for fleets of parallel coding agents with orchestration, policy enforcement, cross-agent review, and observability.

    Use caseUse it to assign work, police permissions, and run independent reviews across Claude Code or Codex adapters in long-running fleets.

    rustcontrol-planefleetpolicy
  17. 17harnessGitHub
    OpenHarness

    maisieyang

    Local-first Python control plane for coding agents featuring Default, Plan, externally judged Goals, sandboxed execution, durable context, plugins, and evals.

    Use caseUse it to drive REPL sessions with independent judges that decide continue, complete, or pause based on artifacts rather than agent claims.

    pythoncontrol-planejudgesandbox
  18. 18harnessGitHub
    HAR

    os-factory

    Open agent harness CLI plus MCP that runs coding agents in isolated worktrees with deterministic launch, verify, and teardown stages.

    Use caseUse it to turn any repo into verified software-factory slots that wrap Claude Code, Cursor, or Codex with project-native checks.

    worktreeverifymcpcli
  19. 19repoGitHub
    Blaxel Sandbox

    blaxel-ai

    Open-source microVM sandbox templates and hub for persistent agent compute with standby resume, MCP tools, filesystem, and process APIs.

    Use caseUse it to give agents dedicated Linux environments that resume in tens of milliseconds while keeping full state across idle periods.

    sandboxmicrovmmcppersistent
  20. 20repoGitHub
    Sandboxer

    hyperterse

    Unified client libraries in Go, Python, and TypeScript that talk directly to E2B, Daytona, Blaxel, Runloop, Fly, or local Docker sandboxes.

    Use caseUse it to switch sandbox providers behind one mental model for create, exec, filesystem, and teardown without rewriting agent code.

    sandboxmulti-providersdkabstraction

02 — In the wild

Articles and releases

  1. releaseTrueFoundry Blog / VentureBeat
    TrueForge open-sources its production agent harness with 30 percent lower cost than Claude Managed Agents

    TrueFoundry

    TrueFoundry released TrueForge under MIT as a vendor-neutral harness with MCP, skills, sandboxing, and compaction. Benchmarks show matching quality at substantially lower token and dollar cost versus managed alternatives on identical tasks.

    releaseharnessopen-sourcecost

    Shipped

  2. articleHacker News
    Show HN: Supafork lets you store, share, and fork agent sessions across harnesses

    supafork

    Supafork provides a GitHub-like layer for AI agent sessions so users can save complete runs from Claude Code, Codex, Pi, and others, organize collections, and import across harnesses. The launch post highlights CLI import plus public or private forking of workflows.

    articlesessionsinteroperabilityshow-hn

    Read

  3. articleHacker News / mouse.dev
    Mouse tops FrontierHarness leaderboard using completion loops on Kimi K3

    Aeroi

    Mouse, built on a customized OpenCode base with specialized rules, skills, and loops, scored highest pass rate and first on time among twelve harnesses all running the same Kimi K3 model. The writeup details overnight Night Shift mode and the completion-loop changes that drove the result.

    articlebenchmarkharnessleaderboard

    Read

Get the daily tools digest

RSS