High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

4 Open Issues Need Help Last updated: Aug 2, 2026

Open Issues Need Help

View All on GitHub
good first issue status:ready type:bug priority:low area:models

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
good first issue status:ready type:test priority:medium area:core

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust
feat: need a logo about 1 month ago
help wanted type:enhancement priority:medium area:docs

High-performance LLM/VLM inference runtime and server for Apple Silicon / CUDA devices

Rust