Popular repositories Loading
-
DeepGEMM
DeepGEMM PublicForked from deepseek-ai/DeepGEMM
DeepGEMM: clean and efficient BLAS kernel library on GPU
Cuda
Repositories
- pega-omni Public
OpenAI-compatible speech serving in Rust: streaming TTS and full-duplex voice over GPT-Live. 128 concurrent PersonaPlex sessions on one GPU.
- pegainfer Public
Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
- pegastore Public
Placement-aware immutable large-object cache for AI workloads: multi-slot values across GPU / DRAM / SSD, shared across nodes.
- DeepGEMM Public Forked from deepseek-ai/DeepGEMM
DeepGEMM: clean and efficient BLAS kernel library on GPU
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…