Trending repositories: benchmarking
5 tracked repositories tagged with benchmarking, ordered by stars. Use the topic filters below to narrow further.
5 of 5 repositories
AlexsJones/llmfit
Hundreds of models & providers. One command to find what runs on your hardware.
AI summary: A terminal tool that evaluates hardware to recommend the best local LLMs based on RAM, CPU, and GPU constraints.
31,176developer-toolsRustMITHKUDS/ClawWork
"ClawWork: OpenClaw as Your AI Coworker - 💰 $15K earned in 11 Hours"
AI summary: A real-world economic benchmark where AI agents complete professional tasks to earn income and survive.
8,299ai-mlPythonMITdavebcn87/pi-autoresearch
Autonomous experiment loop extension for pi
AI summary: An autonomous experimentation loop that continually modifies code, runs benchmarks, and persists improvements.
7,520developer-toolsTypeScriptMITanthropics/original_performance_takehome
Anthropic's original performance take-home, now open for you to try!
AI summary: Anthropic's historical performance engineering take-home assignment, demonstrating complex system optimization challenges.
4,089learningPythonuber/ADR
ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.
AI summary: An enterprise security system developed by Uber to monitor, detect, and mitigate risks from AI coding agents.
1,257securityPythonApache-2.0