Scalable datastore for metrics, events, and real-time analytics
arrow-buffer
Counted from the Cargo.toml manifests of the 36 indexed repositories that declare arrow-buffer as a dependency — not download counts. Dependency data last verified 2026-08-13.
Crates that show up unusually often in arrow-buffer projects. The most distinctive pairings rank first — crates these projects use far more than the average indexed Rust project does, not just crates that are popular everywhere. Each percentage is the share of arrow-buffer projects that also use it.
dbt enables data analysts and engineers to transform their data using the same practices that software engineers use to build applications.
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
Data Agent Ready Warehouse : One for Analytics, Search, AI, Python Sandbox. — rebuilt from scratch. Unified architecture on your S3.
Event streaming platform for agentic AI. Continuously ingest, transform, and serve event streams in real time, at scale.
Simple, Elastic-quality search for Postgres
Apache DataFusion SQL Query Engine
Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data versioning. Compatible with Pandas, DuckDB, Polars, Pyarrow, and PyTorch with more integrations coming..
The open-source Observability 2.0 database. One engine for metrics, logs, and traces — replacing Prometheus, Loki & ES.
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
Official Rust implementation of Apache Arrow
Drop-in Apache Spark replacement written in Rust, unifying batch processing, stream processing, and compute-intensive AI workloads.
A native Rust library for Delta Lake, with bindings into Python
An extensible, state-of-the-art framework for columnar compression, and the fastest FOSS columnar file format. Formerly at @spiraldb, now an Incubation Stage project at LFAI&Data, part of the Linux Foundation.
A portable accelerated SQL query, search, and LLM-inference engine, written in Rust, for data-grounded AI apps and agents.
A cloud-native open source distributed time series database with high performance, high compression ratio and high availability.
Tonbo is an embedded database for serverless and edge runtimes.
Apache Iceberg
Scalable graph analytics database powered by a multithreaded, vectorized temporal engine, written in Rust
Analytical database for data-driven Web applications 🪶
A single-node analytical database engine with geospatial as a first-class citizen
GeoArrow in Rust, Python, and JavaScript (WebAssembly) with vectorized geometry operations
Protocol and libraries for sending and receiving OpenTelemetry data using Apache Arrow
Fully Managed, Streaming Ingestion (CDC) into your Lakehouse
SciRS2 - Scientific Computing and AI in Rust., providing SciPy-compatible APIs while leveraging Rust's performance, safety, and concurrency features. Unlike traditional scientific libraries
The native Rust implementation for Apache Hudi, with C++ & Python API bindings.
A typed, polyglot, functional language
First-class compile‑time Arrow schemas for Rust.
Apache Paimon Rust The rust implementation of Apache Paimon.
Embeddable spreadsheet engine — parse, evaluate & mutate Excel workbooks from Rust, Python, or the browser. Arrow-powered, 320+ functions.
DuckLake took Flight. Welcome to SwanLake.
Graph database native to the cloud. Embedded, multi-tenant, built on object storage.
Geometry and Geography Support for Apache DataFusion
A framework to manage data, continuously
Building block library for using Apache Arrow in Rust WebAssembly modules.
Apache Beam native streaming database built in Rust