Skip to content

Popular repositories Loading

  1. pegainfer pegainfer Public

    Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2

    Rust 712 109

  2. kern kern Public

    A model-agnostic GPU runtime for shipping models as verified programs.

    Rust 48 4

  3. pega-omni pega-omni Public

    OpenAI-compatible speech serving in Rust: streaming TTS and full-duplex voice over GPT-Live. 128 concurrent PersonaPlex sessions on one GPU.

    Rust 28 2

  4. website website Public

    Documentation website for openinfer — Astro Starlight on Cloudflare Workers

    MDX 4 1

  5. pegastore pegastore Public

    Placement-aware immutable large-object cache for AI workloads: multi-slot values across GPU / DRAM / SSD, shared across nodes.

    Rust 4

  6. DeepGEMM DeepGEMM Public

    Forked from deepseek-ai/DeepGEMM

    DeepGEMM: clean and efficient BLAS kernel library on GPU

    Cuda

Repositories

Showing 6 of 6 repositories
  • pega-omni Public

    OpenAI-compatible speech serving in Rust: streaming TTS and full-duplex voice over GPT-Live. 128 concurrent PersonaPlex sessions on one GPU.

    pegainfer-project/pega-omni's past year of commit activity
    Rust 28 2 2 2 Updated Sep 26, 2026
  • kern Public

    A model-agnostic GPU runtime for shipping models as verified programs.

    pegainfer-project/kern's past year of commit activity
    Rust 48 Apache-2.0 4 3 2 Updated Sep 26, 2026
  • pegainfer Public

    Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2

    pegainfer-project/pegainfer's past year of commit activity
    Rust 712 Apache-2.0 109 50 (17 issues need help) 18 Updated Sep 26, 2026
  • website Public

    Documentation website for openinfer — Astro Starlight on Cloudflare Workers

    pegainfer-project/website's past year of commit activity
    MDX 4 1 2 0 Updated Sep 23, 2026
  • pegastore Public

    Placement-aware immutable large-object cache for AI workloads: multi-slot values across GPU / DRAM / SSD, shared across nodes.

    pegainfer-project/pegastore's past year of commit activity
    Rust 4 Apache-2.0 0 0 0 Updated Aug 31, 2026
  • DeepGEMM Public Forked from deepseek-ai/DeepGEMM

    DeepGEMM: clean and efficient BLAS kernel library on GPU

    pegainfer-project/DeepGEMM's past year of commit activity
    Cuda 0 MIT 1,284 0 0 Updated Aug 13, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Rust Cuda MDX