Trending repositories: inference
7 tracked repositories tagged with inference, ordered by stars. Use the topic filters below to narrow further.
7 of 7 repositories
openclaw/openclaw
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
AI summary: A locally hosted, privacy-first personal AI assistant framework supporting massive local deployments.
385,452ai-mlTypeScriptOtherggml-org/llama.cpp
LLM inference in C/C++
AI summary: A highly optimized C/C++ port of LLaMA, designed for running large language models locally on consumer hardware.
122,993ai-mlC++MITmicrosoft/BitNet
Official inference framework for 1-bit LLMs
AI summary: An official inference framework optimized for running 1-bit Large Language Models efficiently on CPUs.
39,816ai-mlC++MITJustVugg/colibri
Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦
AI summary: A minimal C engine for running frontier Mixture of Experts models locally.
23,027ai-mlCApache-2.0NVIDIA/NemoClaw
Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference
AI summary: A high-performance toolkit for fine-tuning and deploying large language models on NVIDIA GPUs using NeMo.
22,079ai-mlTypeScriptApache-2.0Andyyyy64/whichllm
Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.
AI summary: CLI tool to auto-detect hardware and rank best-fitting local LLMs from HuggingFace.
6,152developer-toolsPythonMITantirez/iris.c
Flux 2 image generation model pure C inference
AI summary: A pure C inference pipeline for FLUX.2 image generation models with zero dependencies.
1,967ai-mlCMIT