GitHub
Arbor
观察池 · 暂无正式排名
该项目当前未满足“两个有效维度 + 两种数据源”的主榜门槛。
— 未排名
可观测采用度缺失 —
动量当前有效 · 2026-09-12 12
关注度当前有效 · 2026-09-12 2
信号可信度 依据当前有数据的独立评分维度数量计算。
中2/3 · 1 种数据源
项目介绍
The autonomous research agent that beats Claude Code and Codex by 2.5× on the same compute budget
Give Arbor a benchmark and a goal. It proposes hypotheses, edits code, runs real experiments, and keeps only the gains that survive held-out data — growing a hypothesis tree instead of forgetting what failed.
▶️ Try it in 30 seconds — no API key, no config: > > > Or watch it right now in your browser — nothing to install: ▶️ Live Demo.
and a metric to measure, from model training to harness engineering to data synthesis.
results, failure modes, and distilled insights in the Idea Tree and propagates them upward, so later ideas start smarter instead of scrolling off.
held-out test split, and…
各数据源
1.1k Star
- Star 1.1k
- Fork 126
- 提交 234
- 发布 4