click6067-ship-it

click6067-ship-it/fitllm-engine

Will this LLM fit your GPU or Mac? npx fitllm — accurate memory math for MLA/sliding-window/hybrid/MoE architectures that naive VRAM calculators get wrong by up to 18x. Single file, zero deps, conformance-vector tested. MIT.

2 good first / help-wanted issues · JavaScript · last activity Sep 2, 2026

8 stars 2 forks 8 watchers JavaScript MIT License
amd apple-silicon cli gguf inference kv-cache llama-cpp llm local-llm localllama memory-calculator mla mlx moe nvidia ollama quantization vram vram-calculator will-it-run
2 Open Issues Need Help Last updated: Sep 2, 2026

Open Issues Need Help

View All on GitHub
click6067-ship-it/fitllm-engine
8

Will this LLM fit your GPU or Mac? npx fitllm — accurate memory math for MLA/sliding-window/hybrid/MoE architectures that naive VRAM calculators get wrong by up to 18x. Single file, zero deps, conformance-vector tested. MIT.

JavaScript
#amd#apple-silicon#cli#gguf#inference#kv-cache#llama-cpp#llm#local-llm#localllama#memory-calculator#mla#mlx#moe#nvidia#ollama#quantization#vram#vram-calculator#will-it-run
help wanted good first issue measurement
click6067-ship-it/fitllm-engine
8

Will this LLM fit your GPU or Mac? npx fitllm — accurate memory math for MLA/sliding-window/hybrid/MoE architectures that naive VRAM calculators get wrong by up to 18x. Single file, zero deps, conformance-vector tested. MIT.

JavaScript
#amd#apple-silicon#cli#gguf#inference#kv-cache#llama-cpp#llm#local-llm#localllama#memory-calculator#mla#mlx#moe#nvidia#ollama#quantization#vram#vram-calculator#will-it-run