Try and measure tokio task overhead
- Stars
- 11
- Forks
- 3
- Pushed
- 4y ago
Current GitHub adoption and maintenance signals.
Try and measure tokio task overhead
Demonstration of how to use the rust object_store crate https://crates.io/crates/object_store
No repository description provided.
Test / Demonstration program for parquet footer parsing
VSQL specialization for emacs SQL mode
Scripts for benchmarking https://github.com/apache/arrow-datafusion
Write Parquet with many heterogeneous columns efficiently, with Rust
C++ interface for producer/consumer streams that use coroutines (rather than threads)
Tool to test loading influxdb_iox with high throughput, low cardinality data
Test code for performance analysis
Rust object_store crate
Apache Arrow DataFusion and Ballista query engines
TPC-H benchmark data generation in pure Rust
TPC-DS Data Generator and Data
Useful resources for using the Parquet format
Terminal based, extensible, interactive data analysis tool using SQL
🏋️ Exercise your Parquet readers!
Re run queries that were run on production systems against a local copy of that data for analysis / optimization
CPU bound benchmark for DataFusion: https://github.com/apache/arrow-datafusion
Apache Parquet
Native Rust TPCH support for Datafusion using tpchgen
Comparison tool for influx storage gRPC logs
Benchmark for Adaptive Lossless Floating Point (ALP) in Apache Parquet
Apache DataFusion Web Site
Official Rust implementation of Apache Arrow
Auxiliary files for compatibility and integration tests for Apache Parquet
Apache Parquet
Apache Arrow is a cross-language development platform for in-memory data. It specifies a standardized language-independent columnar memory format for flat and hierarchical data, organized for efficient analytic operations on modern hardware. It also provides computational libraries and zero-copy streaming messaging and interprocess communication. Languages currently supported include C, C++, Java, JavaScript, Python, and Ruby.
Prototype parquet indexing
Adventures in floating point for Apache Parquet ALP encoding
ClickBench: a Benchmark For Analytical Databases
Library for bringing distributed capabilities to Apache DataFusion
AI bot skills for my workflows
Extensible SQL Lexer and Parser for Rust
Mirror of Apache Arrow site
Apache DataFusion SQL Query Engine Testing
A native Rust library for Delta Lake, with bindings into Python
Auto-generate serde implementations for prost types
Analysis of filter pushdown patterns for DataFusion running ClickBench
LakeSail's computation framework with a mission to unify batch processing, stream processing, and compute-intensive AI workloads.
No repository description provided.
A native Delta implementation for integration with any query engine
Attempt to simply reproducer DataFusion IO starvation
Apache DataFusion Python Bindings
No repository description provided.
Apache Iceberg
No repository description provided.
Apache Arrow DataFusion Comet Spark Accelerator
Fast numeric to- and from-string conversion routines.
Zero-cost asynchronous programming in Rust
JSON support for DataFusion (unofficial)
No repository description provided.
Apache Parquet
Demonstration of doing fulltext search on data in parquet files
Convert sequences of Rust objects to Arrow tables
transmute-free Rust library to work with the Arrow format
No repository description provided.
A Rust CPU profiler implemented with the help of backtrace-rs
Apache Arrow Ballista Distributed Query Engine
No repository description provided.
My emacs aglomeration
Arbitrary precision decimal crate for Rust
No repository description provided.
Sqllogictest parser and runner in Rust.
No repository description provided.
No repository description provided.
Apache Arrow Cookbook
Security advisory database for Rust crates published through crates.io
Blazing-fast query execution engine speaks Apache Spark language and has Arrow-DataFusion at its core.
heap profiler for rust
Testing repository for github actions
Query Scheduling Experiments
Parquet Implementation Compare Tool
Dump parquet files from the command line
Scalable datastore for metrics, events, and real-time analytics
Test what happens when cargo test blows up with OOM on github actions
Emacs :heart: Debug Adapter Protocol
InfluxDB/IOx Client for Go
A runtime for writing reliable asynchronous applications with Rust. Provides I/O, networking, scheduling, timers, ...
Apache Parquet format for Rust, hosting the Thrift definition file and the generated .rs file
Empowering everyone to build reliable and efficient software.
Easily assign underlying errors into domain-specific errors while adding context
A data structure to efficiently intern, cache and restore strings.
A gentle Rust tutorial
Very simple command line w/ autocomplete