
Arroyo
by Arroyo · Open-source SQL stream processing engine
16.5 — BenchRank score out of 100
Screenshots of Arroyo
Homepage
Overview
Arroyo is a stream processing engine that runs SQL queries over streaming data. It ships as a single binary that runs locally on MacOS or Linux and deploys with Docker or Kubernetes, with connectors including Kafka, Kinesis, Postgres, MySQL, Redis and Delta Lake. It supports time windows, streaming joins, exactly-once processing, a web UI and a REST API.
- Best for
- Data teams that want real-time streaming pipelines written in SQL without a dedicated streaming team
- Pricing
- No prices are shown on the captured pages; the engine is open source under the Apache 2.0 licence and self-hostable.
- Runs on
- Self-hosted
Strengths and trade-offs
Strengths
- Pipelines are written in standard analytical SQL
- Single binary; runs locally or via Docker and Kubernetes
- Exactly-once processing, time windows and streaming joins
- Apache 2.0 licensed and self-hostable
Trade-offs
- UDFs must be written in Rust; Python is listed as coming soon
- No hosted or managed offering is shown; you run and operate it yourself
- Now owned by Cloudflare, so the roadmap may follow that platform
- Performance and scale claims come from the vendor's own pages only
How this score is made up
Each dimension is scored out of 100 and combined into the headline score using fixed weights.
- MCP support
- 0 out of 100
- Documentation
- 0 out of 100
- Agent friendliness
- 35 out of 100
- Changelog
- 55 out of 100
- Marketing site structure
- 25 out of 100
- Operational trust
- 0 out of 100
Measured, but not part of the score
Useful to know, but not a mark for or against the product — so these do not affect the ranking.
- Openness
- 60 out of 100
- Maintenance
- 100 out of 100
This doesn’t look right — report a problem with Arroyo’s score
Alternatives in Data Pipelines & ETL
Ranked 1
83.3 — BenchRank score out of 100CloudQuery
CloudQuery · Multi-cloud asset inventory with SQL policies and automation
Best for: Platform, security and FinOps teams that need a queryable inventory of a multi-cloud estate.
Ranked 2
77.8 — BenchRank score out of 100OpenSERP
OpenSERP · Open-source SERP API with an optional managed cloud
Best for: Developers and SEO teams needing programmatic multi-engine search data for AI grounding or rank tracking
Ranked 3
75.7 — BenchRank score out of 100Databend
Databend · Open-source Rust data warehouse for SQL, search and vector workloads
Best for: Data teams wanting SQL analytics, full-text and vector search in one warehouse over object storage


