Tuesday, July 14, 2026
HomeBig DataApache Pulsar Demonstrates Finest-in-Class Cloud-Native Value-Efficiency

Apache Pulsar Demonstrates Finest-in-Class Cloud-Native Value-Efficiency

[ad_1]

We’re seeing growing numbers of enterprise initiatives the place knowledge is produced, consumed, analyzed, and reacted to in real-time. On this method, the expertise turns into conscious of what’s occurring inside and round it—making pragmatic, tactical choices by itself. We see this being performed out in transportation, telephony, healthcare, safety and legislation enforcement, finance, manufacturing, and in most sectors of each trade.

Previous to this evolution, the analytical ramifications inherent within the knowledge have been derived lengthy after the occasion that produced or created the info had handed. Now we are able to use expertise to seize, analyze, and take motion based mostly on what’s going on within the second.

This class of knowledge is thought by a number of names: streaming, messaging, reside feeds, real-time, and event-driven. Within the streaming knowledge and message queuing expertise house, there are a variety of standard applied sciences in use, together with Apache Kafka and Apache Pulsar ™.

In January, DataStax, identified for its business help, software program, and cloud database-as-a-service for Apache Cassandra™, launched a brand new line of enterprise for knowledge streaming known as Luna Streaming. DataStax Luna Streaming is a subscription service based mostly on open-source Apache Pulsar. In April, DataStax launched a non-public beta for streaming Pulsar as a service to focus on knowledge engineers, software program engineers, and enterprise architects.

We not too long ago ran a efficiency take a look at evaluating Luna Streaming (Pulsar) and Kafka clusters with Kubernetes. We wished to see if the inherent architectural advantages of Pulsar (tiered storage, decoupled compute and storage, multitenancy) enabled an environment friendly structure that yields tangible efficiency advantages in real-world situations.

We deployed a Kubernetes cluster onto Amazon Net Companies EC2 cases and used the OpenMessaging Benchmark (OMB) take a look at harness to conduct our analysis. We labored with the Confluent fork of the OpenMessaging Benchmark on GitHub. We additionally used the identical {hardware} configuration occasion sorts for Kafka brokers and to co-locate the Pulsar brokers and Bookkeeper nodes to benefit from the 2 massive (2.5TB), quick, locally-attached NVMe solid-state drives.

For Kafka, we spanned the persistent quantity storage throughout each disks. For Pulsar, we created persistent volumes and used each of the native drives for the Bookkeeper ledger and the opposite for the ranges. For the Bookkeeper journal, we provisioned a 100GB gp3 AWS Elastic Block Storage (EBS) quantity with 4,000 IOPS and 1,000 MB/s throughput. Apart from making the most of this storage configuration for each platforms, we carried out no different particular tuning of both platform and most well-liked as a substitute to go along with their “out-of-the-box” configurations as they have been deployed by way of their respective Docker photographs and Helm charts.

Our efficiency testing revealed Luna Streaming had the next common throughput in all of the OMB testing workloads we carried out. By way of dealer node equivalence, we discovered:

3 Luna Streaming nodes @ 5 Kafka nodes

6 Luna Streaming nodes @ 8 Kafka nodes

9 Luna Streaming nodes @ 14 Kafka nodes

We assumed easy linear progress of an enterprise’s streaming knowledge wants over a three-year interval—a “small” cluster (3x Luna Streaming or 5x Kafka) in 12 months 1, a “medium” in 12 months 2 (6x Luna Streaming or 8x Kafka), and a “massive” (9x Luna Streaming or 14x Kafka) in 12 months 3. Utilizing the node equivalences present in our testing above, this could lead to a 33% financial savings in infrastructure prices through the use of Luna Streaming as a substitute of Kafka.

On this state of affairs centered on “peak interval” workloads, we discovered a financial savings of round 50%, relying on the proportion of time the height intervals final.

For our third price state of affairs, we centered on initiatives which will have vital complexity however restricted uncooked throughput necessities, leading to an organizational atmosphere that mandates a excessive variety of matters and partitions to deal with the wide selection of wants throughout the whole enterprise. On this case, we discovered infrastructure financial savings of 75% utilizing Luna Streaming over Kafka.

You’ll be able to obtain the report, with a whole description of the checks and implications of the outcomes, right here.



[ad_2]

RELATED ARTICLES

LEAVE A REPLY

Please enter your comment!
Please enter your name here

Most Popular

Recent Comments