---
title: Blog posts tagged "big_data"
description: Canonical makes open source secure, reliable and easy to use, providing
  support for Ubuntu and a portfolio of enterprise-grade technologies. Founded in
  2004, Canonical operates globally with team members in over 80 countries.
url: https://canonical.com/blog/tag/big_data?format=md
---

# Blog posts tagged "big\_data"

#### [Apache Spark 4.0 beta release – try it now](https://canonical.com/blog/apache-spark-4-0-beta-release-try-it-now)

Apache Spark is a popular framework for developing distributed, parallel data processing applications. Our solution for Apache Spark on Kubernetes has made significant progress in the past year since we launched, adding support for Apache Iceberg, a new GPU accelerated image using the NVIDIA Spark-RAPIDS plugin, and support for the Volcan

---

Data Platform

#### [Deploying and scaling Apache Spark on Amazon AWS EKS](https://canonical.com/blog/deploying-and-scaling-apache-spark-on-amazon-eks)

Move over Hadoop, it’s time for Spark on Kubernetes Apache Spark, a framework for parallel distributed data processing, has become a popular choice for building streaming applications, data lake houses and big data extract-transform-load data processing (ETL). It is horizontally scalable, fault-tolerant, and performs well at high scale. H

---

Data Platform
