Warp is an agentic development environment, born out of the terminal.
tokenizers
Counted from the Cargo.toml manifests of the 136 indexed repositories that declare tokenizers as a dependency — not download counts. Dependency data last verified 2026-08-13.
Crates that show up unusually often in tokenizers projects. The most distinctive pairings rank first — crates these projects use far more than the average indexed Rust project does, not just crates that are popular everywhere. Each percentage is the share of tokenizers projects that also use it.
A lightning-fast search engine API bringing AI-powered hybrid search to your sites and applications.
an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM
YC (S26) | AI that knows what you've seen, said, or heard. Records everything you do, say, hear 24/7, local, private, secure
Minimalist ML framework for Rust
Open-source developer platform to power your entire infra and turn scripts into webhooks, workflows and UIs. Fastest workflow engine (13x vs Airflow). Open-source alternative to Retool and Temporal.
Coding Agent Harness
Memory layer for AI Agents. Replace complex RAG pipelines with a serverless, single-file memory layer. Give your agents instant retrieval and long-term memory.
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured information from PDFs, Office documents, images, and 97+ formats. Available for Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, R, C, TypeScript (Node/Bun/Wasm/Deno)- or use via CLI, REST API, or MCP server.
A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured information from PDFs, Office documents, images, and 97+ formats. Available for Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, R, C, TypeScript (Node/Bun/Wasm/Deno)- or use via CLI, REST API, or MCP server.
Open source Granola AI Alternative
A Datacenter Scale Distributed Inference Serving Framework
Fast, flexible LLM inference
Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data versioning. Compatible with Pandas, DuckDB, Polars, Pyarrow, and PyTorch with more integrations coming..
Spin is the open source developer tool for building and running serverless applications powered by WebAssembly.
ML-powered manga translator, written in Rust.
A blazing fast inference solution for text embeddings models
RuVector is a High Performance, Real-Time, Self-Learning Ai, Vector GNN, Memory DB built in Rust.
A portable accelerated SQL query, search, and LLM-inference engine, written in Rust, for data-grounded AI apps and agents.
Tiny, no-nonsense, self-contained, Tensorflow and ONNX inference
Inference at the speed of light.
Instant, controllable, local pre-trained AI models in Rust
✨ Agentic chat experience in your terminal. Build applications using natural language.
Local first semantic and hybrid BM25 grep / search tool for use by AI and humans!
A high-performance inference engine for AI models
Self hosted, easy to install end to end encrypted storage drive
A Modern Embedded SQL Database written in Rust
Voice-to-text with push-to-talk for Wayland compositors
NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.
Rust library for generating vector embeddings, reranking locally!
Flow-Like: Strongly Typed Enterprise Scale Workflows. Built for scalability, speed, seamless AI integration and rich customization.
The fastest PDF library for Python and Rust. Text extraction, image extraction, markdown conversion, PDF creation & editing. 0.8ms mean, 5× faster than industry leaders, 100% pass rate on 3,830 PDFs. MIT/Apache-2.0.
rvLLM: High-performance LLM inference in Rust. Drop-in vLLM replacement.
A multi-agent framework written in Rust that enables you to build, deploy, and coordinate multiple intelligent agents
Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.
The open source Unity Dev Agent
Japanese Input Method System for Linux, macOS, Neural Kana-Kanji Conversion Engine
Run full Kimi K3 on a single device. And an OpenAI-compatible API server for local chat and coding agents.
🦀 Low-level 3D Computer Vision library in Rust
Pure Rust Inference Engine
Own your PaaS
Split text into semantic chunks, up to a desired chunk size. Supports calculating length by characters and tokens, and is callable from Rust and Python.
Rust Agent Development Kit (ADK-Rust): Build AI agents in Rust with modular components for models, tools, memory, realtime voice, and more. ADK-Rust is a flexible framework for developing AI agents with simplicity and power. Model-agnostic, deployment-agnostic, optimized for frontier AI models. Includes support for real-time voice agents.
NextPlaid, ColGREP: Multi-vector search, from database to coding agents.
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler and cleaner..
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.