#23 · Primary category: Business Intelligence & Analytics
starrocks
The world's fastest open-source query engine for sub-second analytics on and off the data lakehouse.
Project last updated:08/29/26
GitHub Stars
12.1K
Forks
2.6K
Contributors
680
License
Apache-2.0
Why we included this project
Data teams that need sub-second answers over very large analytical datasets will find a serious workhorse here. StarRocks is a distributed, vectorized SQL engine built on MPP and columnar storage, a combination that suits multi-dimensional reporting, real-time updates, and ad-hoc exploration. Its external catalog layer reads Hive, Iceberg, Hudi, or Delta Lake tables in place, so you can run fast analytics on the lake without migrating data into a proprietary store. Asynchronous materialized views and caching speed up repeated dashboard queries against those external sources, while native tables absorb high-ingest real-time workloads. Whether you want a faster compute layer over an existing data lake or a dedicated OLAP store, this covers both without forcing you to pick a side.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
spark
Apache Spark - A unified analytics engine for large-scale data processing
metabase
The easy-to-use open source Business Intelligence and Embedded Analytics tool that lets everyone work with data :bar_chart:
streamlit
Streamlit — A faster way to build and share data apps.
duckdb
DuckDB is an analytical in-process SQL database management system
ToolJet
ToolJet is the open-source foundation of ToolJet AI - the enterprise app generation platform for building internal tools, dashboard, business applications, workflows and AI agents 🚀