Popular repositories Loading
Repositories
Showing 3 of 3 repositories
- ferrumox.com Public
- fox Public
High-performance LLM inference engine — drop-in replacement for Ollama with faster multi-turn inference, lower TTFT, and higher throughput through prefix caching and continuous batching.
- rabbit Public
Run Kimi K3 (2.8T MoE) on a 128GB-RAM consumer machine in pure Rust, experts streamed from disk. Tiny engine, immense model.
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…