Trending repositories: inference

7 tracked repositories tagged with inference, ordered by stars. Use the topic filters below to narrow further.

Filter by topic

7 of 7 repositories

  • openclaw/openclaw

    Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

    AI summary: A locally hosted, privacy-first personal AI assistant framework supporting massive local deployments.

    385,452ai-mlTypeScriptOther
  • ggml-org/llama.cpp

    LLM inference in C/C++

    AI summary: A highly optimized C/C++ port of LLaMA, designed for running large language models locally on consumer hardware.

    122,993ai-mlC++MIT
  • microsoft/BitNet

    Official inference framework for 1-bit LLMs

    AI summary: An official inference framework optimized for running 1-bit Large Language Models efficiently on CPUs.

    39,816ai-mlC++MIT
  • JustVugg/colibri

    Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

    AI summary: A minimal C engine for running frontier Mixture of Experts models locally.

    23,027ai-mlCApache-2.0
  • NVIDIA/NemoClaw

    Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

    AI summary: A high-performance toolkit for fine-tuning and deploying large language models on NVIDIA GPUs using NeMo.

    22,079ai-mlTypeScriptApache-2.0
  • Andyyyy64/whichllm

    Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.

    AI summary: CLI tool to auto-detect hardware and rank best-fitting local LLMs from HuggingFace.

    6,152developer-toolsPythonMIT
  • antirez/iris.c

    Flux 2 image generation model pure C inference

    AI summary: A pure C inference pipeline for FLUX.2 image generation models with zero dependencies.

    1,967ai-mlCMIT