A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.
hf-hub
Counted from the Cargo.toml manifests of the 86 indexed repositories that declare hf-hub as a dependency — not download counts. Dependency data last verified 2026-08-13.
Crates that show up unusually often in hf-hub projects. The most distinctive pairings rank first — crates these projects use far more than the average indexed Rust project does, not just crates that are popular everywhere. Each percentage is the share of hf-hub projects that also use it.
an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM
A free, open source, and extensible speech-to-text application that works completely offline.
Search infrastructure for AI
YC (S26) | AI that knows what you've seen, said, or heard. Records everything you do, say, hear 24/7, local, private, secure
Minimalist ML framework for Rust
Open-source developer platform to power your entire infra and turn scripts into webhooks, workflows and UIs. Fastest workflow engine (13x vs Airflow). Open-source alternative to Retool and Temporal.
Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
💥 Fast State-of-the-Art Tokenizers optimized for Research and Production
A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured information from PDFs, Office documents, images, and 97+ formats. Available for Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, R, C, TypeScript (Node/Bun/Wasm/Deno)- or use via CLI, REST API, or MCP server.
A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured information from PDFs, Office documents, images, and 97+ formats. Available for Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, R, C, TypeScript (Node/Bun/Wasm/Deno)- or use via CLI, REST API, or MCP server.
Open source Granola AI Alternative
A Datacenter Scale Distributed Inference Serving Framework
Fast, flexible LLM inference
Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data versioning. Compatible with Pandas, DuckDB, Polars, Pyarrow, and PyTorch with more integrations coming..
ML-powered manga translator, written in Rust.
A blazing fast inference solution for text embeddings models
RuVector is a High Performance, Real-Time, Self-Learning Ai, Vector GNN, Memory DB built in Rust.
Drop-in Apache Spark replacement written in Rust, unifying batch processing, stream processing, and compute-intensive AI workloads.
Distributed AI/LLM for the people. Share compute privately or publicly to power your agents and chat.
A portable accelerated SQL query, search, and LLM-inference engine, written in Rust, for data-grounded AI apps and agents.
Inference at the speed of light.
Local first semantic and hybrid BM25 grep / search tool for use by AI and humans!
Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale 🏓🦙 Alternative to projects like llm-d, Docker Model Runner, etc but with less moving parts and simple deployments built around ggml ecosystem. Runs on CPU and GPU.
A Modern Embedded SQL Database written in Rust
Rust library for generating vector embeddings, reranking locally!
A multi-agent framework written in Rust that enables you to build, deploy, and coordinate multiple intelligent agents
Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.
The open source Unity Dev Agent
Japanese Input Method System for Linux, macOS, Neural Kana-Kanji Conversion Engine
🦀 Low-level 3D Computer Vision library in Rust
Rust Agent Development Kit (ADK-Rust): Build AI agents in Rust with modular components for models, tools, memory, realtime voice, and more. ADK-Rust is a flexible framework for developing AI agents with simplicity and power. Model-agnostic, deployment-agnostic, optimized for frontier AI models. Includes support for real-time voice agents.
NextPlaid, ColGREP: Multi-vector search, from database to coding agents.
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler and cleaner..
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.
Build apps powered by on-device AI
Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.
Models and examples built with Burn
ANOLISA - Agentic Nexus Operating Layer & Interface System Architecture
A comprehensive Rust translation of the code from Sebastian Raschka's Build an LLM from Scratch book.
Open coding agent, provider agnositc.
Blazing-fast LLM inference in pure Rust. No PyTorch and Python runtime.
MoFA - Modular Framework for Agents. Modular, Compositional and Programmable.
Give your agent a proper IDE and OS. The sensorimotor cortex for coding agents (OpenCode + Pi), part of CortexKit: symbol-aware edits, semantic search, code health, fast grep/glob, bash compression, background tasks, PTY.