Minimalist ML framework for Rust
cudarc
Counted from the Cargo.toml manifests of the 37 indexed repositories that declare cudarc as a dependency — not download counts. Dependency data last verified 2026-08-13.
Crates that show up unusually often in cudarc projects. The most distinctive pairings rank first — crates these projects use far more than the average indexed Rust project does, not just crates that are popular everywhere. Each percentage is the share of cudarc projects that also use it.
A Datacenter Scale Distributed Inference Serving Framework
ML-powered manga translator, written in Rust.
A blazing fast inference solution for text embeddings models
An extensible, state-of-the-art framework for columnar compression, and the fastest FOSS columnar file format. Formerly at @spiraldb, now an Incubation Stage project at LFAI&Data, part of the Linux Foundation.
A portable accelerated SQL query, search, and LLM-inference engine, written in Rust, for data-grounded AI apps and agents.
Tiny, no-nonsense, self-contained, Tensorflow and ONNX inference
Inference at the speed of light.
Multi-platform high-performance compute language extension for Rust.
Apache Mahout - an environment for quickly creating scalable, performant machine learning applications.
rvLLM: High-performance LLM inference in Rust. Drop-in vLLM replacement.
🦀 Low-level 3D Computer Vision library in Rust
Pure Rust + CUDA LLM inference engine
Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
Pure Rust Inference Engine
NextPlaid, ColGREP: Multi-vector search, from database to coding agents.
From-scratch Rust+CUDA inference engine, bit-exact by construction — NVFP4, MoE, MTP speculative decoding, tuned against measured limits of one RTX 5090 Laptop (sm_120a).
memra — from-scratch LLM inference engine for NVIDIA RTX 50-series (Rust + CUDA)
SciRS2 - Scientific Computing and AI in Rust., providing SciPy-compatible APIs while leveraging Rust's performance, safety, and concurrency features. Unlike traditional scientific libraries
High-performance KV cache storage for LLM inference — GPU offloading, SSD caching, and cross-node sharing via RDMA. Works with vLLM and SGLang.
KV cache control plane and evaluation platform for LLM inference
Rust-native GPU operator library for LLM inference, built with cuda-oxide
Camelid: a Rust-native local inference backend with evidence-gated model compatibility.
High-performance quantitative finance in Rust — 120+ stochastic processes, option pricing, calibration, fixed income, risk & copulas, with SIMD/GPU acceleration and Python bindings.
Rust inference package experiments
Accelerated Zero-knowledge Virtual Machine by Non-uniform Prover Based on GKR Protocol
Protein and molecule viewer, editor, simulator
ODE solver library in Rust
A machine learning library built with user convenience at its core
Next Generation Machine Learning, Statistics and Deep Learning in PURE Rust
A general-purpose extensible tensor computation library in Rust with CPU/GPU support
A Rust Library for High-Performance Tensor Exchange with Python
iris-mpc repository
Rust Implementation for WebNN
Minimalist, performant and auditable verifiable RISC-V vm written in Rust