Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
Realistic first-contribution target: active project, 143 open issues, small enough for a newcomer PR to land.
Derived from this repo's Cargo dependencies and activity — difficulty is an estimate from codebase size and scope.
Matched by dependency overlap in Cargo manifests — each card notes the most distinctive crates both projects share.
Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang.
A DuckDB extension for graph data analytics
The official Rust library for the Edgee AI Gateway
Rust crate for Beacon Atlas relay — AI agent registration, heartbeat, SEO-enhanced discoverability. Published on crates.io.
Mock REST APIs from JSON with zero coding within seconds.
mDNS, SSDP/DIAL, WSD, WoL reflector
Rust crate package to link to a system libz (zlib)
A Datacenter Scale Distributed Inference Serving Framework
High-performance LLM inference engine — drop-in replacement for Ollama with faster multi-turn inference, lower TTFT, and higher throughput through prefix caching and continuous batching.