Cloudflare Launches Serverless Analytics Platform Basin, Challenging Snowflake and Databricks
Cloudflare has introduced Cloudflare Basin, a serverless data platform built on Apache Iceberg and its egress-free R2 storage, aiming to lower the cost of analytics infrastructure and move into a market led by Snowflake and Databricks.
Basin runs on Cloudflare's global network and is built on the open-source Apache Iceberg table format, designed for petabyte-scale analytics datasets, together with Cloudflare R2, the company's distributed object storage service that carries no egress fees. Cloudflare says the combination removes the need for customers to invest in dedicated servers or pay to move data between clouds and systems.
Analytics workloads have traditionally required a team of data engineers to set up, manage and maintain the server systems that store and organize data, plus the skills to connect separate systems into pipelines that feed an analytics engine. Cloudflare says those operations are prohibitively expensive, citing egress fees charged to move data out of public clouds along with salaries and system overhead, costs that have generally limited sophisticated analytics to the largest organizations.
Because Iceberg allows engines such as Apache Spark and DuckDB to read data directly at its source, customers do not need to copy information, move files or convert formats. Data is left in place, and only the insights produced by querying it are extracted, with data collection, preparation and analysis running on Cloudflare's network. Pricing is based on how much data a customer analyzes.
Michael Ni of Constellation Research pointed to the economics of the Iceberg data stack, which makes storage and compute more interchangeable. "Cloudflare already has the global infrastructure, serverless compute and an egress-free storage model," he said. "And it already sits in the path of a lot of application, log and event data, so that reduces its data movement costs significantly."
Cloudflare Chief Technology Officer Dane Knecht said developers should not have to learn how to operate data infrastructure in order to query their own information. "With Cloudflare Basin, we are bringing the same serverless model that developers expect from Cloudflare to analytics," he said. "No clusters to manage, no unnecessary data movement and open standards that keep customers in control of their data."
The launch puts Cloudflare into competition with established data warehouse vendors while it expands beyond the content delivery and security products it is best known for. Cloudflare says Basin is not a blanket replacement for every data warehouse workload, but argues that many analytics jobs are too small to justify dedicated server clusters and the cost of moving data around.
Ni said he does not expect larger companies to abandon Databricks or Snowflake soon, but sees an opening for Cloudflare at the edge of the traditional data warehouse market. According to SiliconANGLE, which reported the launch, Cloudflare intends to use its interconnected global network to let companies analyze data closer to where it is created, reducing latency and cost relative to centralized cloud data warehouses.