UiPath/coder_eval
Evaluate & benchmark AI coding agents and Claude Code skills — sandboxed, reproducible YAML eval suites for Claude Code, Codex & Gemini, with A/B experiments and CI gates.
Python
Stars
105
Forks
1
License
Apache-2.0
Open Issues
10
Updated
1h ago
No README available
Related Projects
affaan-m/ECC
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
232.2k
JavaScript
NousResearch/hermes-agent
The agent that grows with you
219.0k
Python
Significant-Gravitas/AutoGPT
AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
185.6k
Python
f/prompts.chat
f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.
166.2k
HTML
firecrawl/firecrawl
The API to search, scrape, and interact with the web at scale. 🔥
154.5k
TypeScript
langgenius/dify
Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without rebuilding the stack.
149.8k
TypeScript
Statistics
- Stars
- 105
- Forks
- 1
- Open Issues
- 10
- Created
- Jul 09, 2026
- Updated
- Jul 23, 2026