π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
safetensors
Counted from the Cargo.toml manifests of the 43 indexed repositories that declare safetensors as a dependency — not download counts. Dependency data last verified 2026-08-13.
Crates that show up unusually often in safetensors projects. The most distinctive pairings rank first — crates these projects use far more than the average indexed Rust project does, not just crates that are popular everywhere. Each percentage is the share of safetensors projects that also use it.
A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.
YC (S26) | AI that knows what you've seen, said, or heard. Records everything you do, say, hear 24/7, local, private, secure
Minimalist ML framework for Rust
Burn is a next generation tensor library and Deep Learning Framework that doesn't compromise on flexibility, efficiency and portability.
Fast, flexible LLM inference
Spin is the open source developer tool for building and running serverless applications powered by WebAssembly.
⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.
Rust bindings for the C++ api of PyTorch.
A blazing fast inference solution for text embeddings models
3D Reconstruction for all
Tiny, no-nonsense, self-contained, Tensorflow and ONNX inference
Inference at the speed of light.
NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.
Rust library for generating vector embeddings, reranking locally!
rvLLM: High-performance LLM inference in Rust. Drop-in vLLM replacement.
Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.
Pure Rust + CUDA LLM inference engine
Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
Pure Rust Inference Engine
Build apps powered by on-device AI
Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.
ONNX neural network inference engine
MLX-based experimental inference engine
Rust MCP server for multi-agent coordination: 34 tools, Git-backed archive, SQLite indexing, advisory file locks, and an interactive TUI console
A Rust-based AI agent orchestrator with Neo4j knowledge graph, Meilisearch semantic search, and Tree-sitter code parsing.
HyprStream: agentic infrastructure for continous online-learning applications
A machine learning library for Python and Rust, for PyTorch, Tensorflow and SKLearn models
Next Generation Machine Learning, Statistics and Deep Learning in PURE Rust
Terminal hypervisor for AI agent swarms: real-time pane capture, state-machine pattern detection, and a JSON API for coordinating fleets of coding agents across WezTerm
High Performance Low-Latency Frameworks
Two-tier hybrid search for Rust: sub-millisecond initial results via potion-128M, quality-refined rankings in 150ms via MiniLM-L6-v2. Combines lexical (Tantivy BM25) and semantic (vector cosine) search with Reciprocal Rank Fusion. Progressive iterator API, f16 SIMD vector index, feature-gated compilation.
Expert Kit is an efficient foundation of Expert Parallelism (EP) for MoE model Inference on heterogenous hardware
rust recursive deep learning framework
pure rust inference engine
Local STT server powered by GigaAM v3.
Local AI image generation CLI — FLUX, SD 1.5, SDXL & Z-Image diffusion models on your GPU
Multi-provider AI agent CLI written in Rust
Make Whisper work with APR model format and run in WASM
sparse ternary AI stack enabling efficient frontier intelligence without hyperscaler-scale infrastructure.
Decentralized AI Infrastructure