Trending repositories: benchmarking

5 tracked repositories tagged with benchmarking, ordered by stars. Use the topic filters below to narrow further.

Filter by topic

5 of 5 repositories

  • AlexsJones/llmfit

    Hundreds of models & providers. One command to find what runs on your hardware.

    AI summary: A terminal tool that evaluates hardware to recommend the best local LLMs based on RAM, CPU, and GPU constraints.

    31,176developer-toolsRustMIT
  • HKUDS/ClawWork

    "ClawWork: OpenClaw as Your AI Coworker - 💰 $15K earned in 11 Hours"

    AI summary: A real-world economic benchmark where AI agents complete professional tasks to earn income and survive.

    8,299ai-mlPythonMIT
  • davebcn87/pi-autoresearch

    Autonomous experiment loop extension for pi

    AI summary: An autonomous experimentation loop that continually modifies code, runs benchmarks, and persists improvements.

    7,520developer-toolsTypeScriptMIT
  • anthropics/original_performance_takehome

    Anthropic's original performance take-home, now open for you to try!

    AI summary: Anthropic's historical performance engineering take-home assignment, demonstrating complex system optimization challenges.

    4,089learningPython
  • uber/ADR

    ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.

    AI summary: An enterprise security system developed by Uber to monitor, detect, and mitigate risks from AI coding agents.

    1,257securityPythonApache-2.0