Skip to content

Popular repositories Loading

  1. fox fox Public

    High-performance LLM inference engine — drop-in replacement for Ollama with faster multi-turn inference, lower TTFT, and higher throughput through prefix caching and continuous batching.

    Rust 185 28

  2. rabbit rabbit Public

    Run Kimi K3 (2.8T MoE) on a 128GB-RAM consumer machine in pure Rust, experts streamed from disk. Tiny engine, immense model.

    Rust 121 15

  3. ferrumox.com ferrumox.com Public

    HTML

Repositories

Showing 3 of 3 repositories

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Loading…

Most used topics

Loading…