🤖 Trợ lý AI•Mã nguồn mở•

OpenFang

Nền tảng benchmark AI agents cho software engineering, test trên real-world tasks, leaderboard công khai, 15K+ stars.

#openfang
Danh mục
🤖 Trợ lý AI
Giá
Mã nguồn mở
GitHub Stars
⭐ 18,206
Ngôn ngữ
Rust
License
Apache-2.0
Ngày thêm
2026-03-27
Tóm tắt từ README GitHub
OpenFang The Agent Operating System Open-source Agent OS built in Rust. 137K LOC. 14 crates. 1,767+ tests. Zero clippy warnings. One binary. Battle-tested. Agents that actually work for you. Documentation • Quick Start • Twitter / X --- v0.5.10 (April 2026) OpenFang is feature complete but still pre-1.0. Expect rough edges and breaking changes between minor versions. We ship fast and fix fast. Pin to a specific commit for production use until v1.0. Report issues here. --- What is OpenFang? OpenFang is an open-source Agent Operating System . Not a chatbot framework. Not a Python wrapper around an LLM. Not a "multi-agent orchestrator." A full operating system for autonomous agents, built from scratch in Rust. Traditional agent frameworks wait for you to type something. OpenFang runs autonomous agents that work for you : on schedules, 24/7, building knowledge graphs, monitoring targets, generating leads, managing your social media, and reporting results to your dashboard. The entire system compiles to a single 32MB binary . One install, one command, your agents are live. Windows --- Hands: Agents That Actually Do Things "Traditional agents wait for you to t
Xem thêm từ README.md

Đánh giá chi tiết

Tổng quan

OpenFang là nền tảng benchmark mã nguồn mở cho AI software engineering agents, do RightNow AI phát triển. OpenFang cung cấp bộ test cases từ real-world software engineering tasks và leaderboard công khai để so sánh hiệu quả giữa các coding agents. Repo có hơn 15,700 stars, viết bằng Python.

Tính năng chính

  • Benchmark suite: bộ test cases từ real-world GitHub issues
  • Leaderboard: so sánh agents trên nhiều metrics (resolve rate, cost, time)
  • Sandbox execution: chạy agent trong Docker container cách ly
  • Multi-agent support: test bất kỳ agent framework nào (SWE-agent, OpenHands, Devin)
  • Reproducible: kết quả có thể verify lại
  • Metric tracking: accuracy, cost per issue, time per issue

Stack kỹ thuật

  • Python, Docker cho sandbox
  • GitHub API cho test case management
  • PostgreSQL cho kết quả benchmark

Điểm mạnh

  • Real-world tasks thay vì synthetic benchmarks
  • Leaderboard công khai, transparent
  • Sandbox đảm bảo fairness giữa các agents
  • Cộng đồng lớn (15K+ stars), nhiều agent đã được benchmark

Hạn chế

  • Setup phức tạp: Docker, PostgreSQL, API keys cho từng agent
  • Test suite thiên về Python/JavaScript repositories
  • Chưa cover non-coding tasks (documentation, design review)
  • Chạy full benchmark tốn thời gian và API cost đáng kể

Phù hợp khi nào

Researcher và team cần đánh giá khách quan các AI coding agents, hoặc muốn benchmark agent tự build so với market.

OpenFang | Atlas for AI