Skip to content
#

long-horizon-agents

Here are 29 public repositories matching this topic...

The long-horizon computer-use harness. Run AI agents across desktop apps and the CLI for extended periods while preserving task state and making reliable progress on complex workflows. Features fresh-context execution, durable verified state, independent auditing, recoverable progress, and native Claude Code / Codex / OpenClaw integration.

  • Updated Aug 10, 2026
  • Python
coder_eval

Test that your Claude Code skills, MCP servers, and CLIs actually work when an agent uses them — sandboxed YAML suites, activation checks, A/B experiments, CI gates.

  • Updated Aug 10, 2026
  • Python
awesome-loop-engineering

🔁 Build reliable recurring AI-agent systems: 956 resources, 22 operational patterns, 22 loop contracts, 8 runtime starters, an interactive atlas, and a structured dataset.

  • Updated Aug 10, 2026
  • Python

EvoX Genesis is an autonomous system for long-horizon software evolution that recursively builds, continues, and transforms complex software from high-level objectives.

  • Updated Aug 10, 2026
  • Elixir

Recoverable long-horizon AI agents — a framework-agnostic reference harness + recovery-faithful live benchmark. Thesis: "Checkpoints Are Compactions" via Re-grounding Recovery. 0.x: v1.0 held until a powered live-LLM study confirms the claims.

  • Updated Jul 2, 2026
  • Python

Server-authoritative Minecraft AI agent with a Python LLM brain and Fabric + Carpet/Scarpet body. Autonomous navigation, gathering, crafting, combat, recovery, persistent memory, governed tools, and long-horizon play.

  • Updated Jul 30, 2026
  • Python

Bounded context for long-horizon LLM agents: carry a small re-grounded signal instead of replaying the transcript, so a parent agent's footprint stays flat as the task grows. Recompute-verifiable; null results published; cheaper, not smarter.

  • Updated Aug 5, 2026
  • Python

A driver/worker/store discipline for bounded-context agentic execution: stateless workers, a born-tiered store, and structural audit independence so complex multi-session LLM work survives without losing the thread. Platform-neutral — the runtime ships as a Claude Code plugin.

  • Updated Aug 1, 2026

Improve this page

Add a description, image, and links to the long-horizon-agents topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the long-horizon-agents topic, visit your repo's landing page and select "manage topics."

Learn more