autonomous agent research and evaluation
Recent window: last 5.2 months (2026-04-24 → 2026-09-28), compared with the prior 5.2 months.
These products offer reinforcement learning frameworks, evaluation benchmarks, and tooling designed to train and stress-test autonomous AI agents. They are used by AI researchers and machine learning engineers developing decision-making models and scientific discovery systems. Unlike generic workflow automation or consumer copilot tools, this cluster centers on recursive self-improvement, policy optimization, and rigorous agent capability testing.
Metrics
- Stage
- crowded
- Recent count
- 149
- Prior count
- 12
- Total count
- 188
- Momentum
- 300.00
- Attention
- 0.15
- Crowding
- 0.77
- Concentration
- 0.85
- Opportunity
- 0.38
Opportunity components
- Attention
- 0.15
- Low crowding
- 0.23
- Momentum (normalized)
- 1.00
- Low concentration
- 0.15
Monthly trajectory
Source split
- github
- 125 (0.84)
- hn
- 7 (0.05)
- ph
- 10 (0.07)
- yc
- 7 (0.05)
Dominant source: github · Divergence: 0.79
Similar themes
- specialized AI models and agent reasoning tools (0.80)
- modular ai agent skills and toolkits (0.76)
- 3d modeling and graphics engines (0.73)
- scientific computing and research algorithms (0.73)
- ai agent infrastructure and tooling (0.67)
- low-level systems and developer tools (0.63)
- AI coding agents and developer tools (0.62)
- desktop AI assistants and agent platforms (0.60)
Members
| Name | Source | Upvotes ▲ | Launched |
|---|---|---|---|
| survival-rl | GITHUB | 12 | 2026-09-27 |
| AlphaResearchOS | GITHUB | 13 | 2026-09-22 |
| Visibl Semiconductors: The First AI-Enabled Coordination Layer for Chip Design | YC | 13 | 2026-02-19 |
| Parallel Agentic Search on the Twitter Algorithm | HN | 13 | 2026-01-20 |
| Biopharma Bench 0.1 – Evaluating AI on Autonomous Drug Development | YC | 14 | 2026-09-25 |
| Fractal | HN | 14 | 2026-07-21 |
| JevAny | GITHUB | 17 | 2026-09-22 |
| multi-agent-relay | GITHUB | 18 | 2026-09-16 |
| Next Move Theory gives you an algorithm for every product decision | HN | 18 | 2026-06-22 |
| algorithm-practice-roadmap | GITHUB | 18 | 2026-09-27 |
| SolveEdit | GITHUB | 18 | 2026-09-28 |
| minesweeper | GITHUB | 18 | 2026-09-21 |
| harness-router | GITHUB | 18 | 2026-09-25 |
| llm-multi-agent-game-evaluation | GITHUB | 19 | 2026-09-23 |
| 10x Science: Unlocking the future of drug development 🔑🧬 | YC | 19 | 2026-02-12 |
| AI-Has-Taste | GITHUB | 19 | 2026-09-12 |
| ptcg-population-rl | GITHUB | 21 | 2026-09-12 |
| WMRL | GITHUB | 21 | 2026-09-10 |
| graph-engineering | GITHUB | 21 | 2026-09-09 |
| Swarmmy | GITHUB | 22 | 2026-09-22 |
| research-solution-supervisor | GITHUB | 22 | 2026-09-09 |
| falsify-the-problem | GITHUB | 25 | 2026-09-10 |
| LoopGain | HN | 31 | 2026-07-15 |
| auto-team | GITHUB | 34 | 2026-09-24 |
| easyplot | GITHUB | 42 | 2026-09-23 |
| ScienceBuddy | GITHUB | 57 | 2026-09-14 |
| Aha-Engine | GITHUB | 63 | 2026-09-10 |
| Agent Swarm | HN | 63 | 2026-02-26 |
| Valori | PH | 64 | 2026-09-22 |
| laborix | GITHUB | 87 | 2026-09-30 |
| Maximem Synap | PH | 134 | 2026-09-24 |
| Agent Skills Leaderboard | HN | 135 | 2026-01-20 |
| Mediator.ai | HN | 160 | 2026-04-20 |
| djev-spark | GITHUB | 172 | 2026-09-18 |
| KLPO | GITHUB | 180 | 2026-09-19 |
| Terranox AI: AI-powered uranium discovery | YC | 218 | 2026-02-25 |
| RSIAgent | GITHUB | 300 | 2026-09-13 |
| MiniMax M2.7 | PH | 379 | 2026-03-19 |