AlexsJones/llmfit
One terminal command tells you which AI models will actually run on your computer
llmfit is a terminal tool that detects your machine's RAM, CPU, and GPU, then ranks hundreds of language models by how well they'd actually run on your hardware. A single command shows a scored table right in your terminal. It also lets you measure real speed on your own machine and share results back so estimates get better for everyone.
What it does
- Supports hundreds of models across multiple local runtime providers, including Ollama, llama.cpp, MLX, Docker Model Runner, and LM Studio
- Detects RAM, CPU, and GPU/VRAM, and scores each model across four dimensions: memory fit, estimated speed, quality, and context length
- Speed estimates use a memory-bandwidth-based formula combined with real measurements collected from the community
- Users can benchmark on their own hardware and submit results as a PR from the interactive TUI, with merged results becoming verified numbers for others with identical hardware in the next release
- Handles multi-GPU setups and MoE (Mixture-of-Experts) architectures, which only activate a subset of parameters
Why it matters
Anyone trying to run AI models locally usually has to guess or trial-and-error their way into finding a model that fits their hardware, and this tool automates that decision. As more real user measurements accumulate, the project's estimates shift from guesses to verified numbers that benefit the whole community.
Terms in this repo
- TUI · A text-based interactive interface you navigate inside a terminal
- MoE (Mixture-of-Experts) · A neural network design that activates only part of its total parameters at once, needing less memory than its full size suggests
- VRAM · Dedicated memory built into a graphics card (GPU)
- tok/s · Tokens processed per second, a speed measure for language models
- TTFT · Time-to-first-token, how long it takes before a model starts producing its first output after a prompt
Repository description (English)
Hundreds of models & providers. One command to find what runs on your hardware.
Open on GitHubTrending repos
- cathrynlavery/diagram-designA skill that makes AI coding tools draw magazine-quality diagrams instead of generic rounded boxes
- public-apis/public-apisA giant crowd-curated directory of free APIs for developers
- semantica-agi/semanticaAn open-source graph infrastructure that lets AI agents show their work, not just their answers
- cactus-compute/needleA 14MB AI model small enough to run tool-calling on a phone or watch, without internet
- unslothai/unslothA desktop app that lets you run and train AI models on your own computer, no coding required
- macro-inc/macroAn all-in-one workspace where email, chat, docs, tasks, and CRM are cross-linked and share one AI memory
- harry0703/MoneyPrinterTurboAn open-source tool that turns a single topic or keyword into a finished short video, complete with script, footage, subtitles, and music
- basecamp/omarchyA ready-made, opinionated Linux setup built by DHH
Latest from METAL LAB
- NVIDIA's 300 Verified Skills Lift Correctness by 41 Points
- Wave your hand at a webcam, hear a theremin: browser instrument released
- Meta AI launches desktop app for Mac, can read an entire app window
- Factory Commits $100M to Partner Network, Pushes to Scale Software Factories
- SpaceX approached Cognition for acquisition four days after closing Cursor deal