Apache Spark 4.0 beta release – try it now

Apache spark 4. 0 beta release – try it now 3

Apache Spark is a popular framework for developing distributed, parallel data processing applications. Our solution for Apache Spark on Kubernetes has made significant progress in the past year since we launched, adding support for Apache Iceberg, a new GPU accelerated image

A data warehouse, on a cloud-native data lake with Apache Kyuubi

We’ve also been busy adding initial support for Apache Kyuubi to our Charmed Spark solution, so that you can deploy an enterprise-grade, fault-tolerant, ANSI-SQL-compliant data warehouse on your Kubernetes data lake infrastructure, building a so-called ‘lakehouse’. You can deploy a comprehensive, hyper-automated data lake infrastructure using our all-open source control plane, software defined storage and cloud-native compute infrastructure solutions. We’ve even built a couple of runbooks that should get you started in both cloud and on-premise contexts.

There are many benefits to adopting the cloud-native approach to building a data lake:

Disaggregated storage means you can scale and manage your storage tier independently of your compute tier, shut down or scale down your compute tier when not being used, and take advantage of cost-optimized object storage systems for hosting your big data.
Using cloud-native technologies for the compute tier (ie. Kubernetes) assures a high level of portability between infrastructure providers, so you need not be locked in to any one cloud service provider or data center systems vendor.
Using a cloud-native approach means you can build clusters both on the cloud and in your on-premise facility, but use the same operational management approach in a consistent way.
You can potentially “bin pack” other applications onto the same cloud-native platform as your data lake infrastructure, for more efficient resource utilization, if you so wish.
Do you want to use GPUs to accelerate your Spark applications? Kubernetes has excellent support for exposing GPUs to Spark and can greatly simplify setup of this useful acceleration feature.

While we do have some work to do until our Kyuubi integration is fully ready for business, you can already try it out – see our docs for the lowdown.

Spark 4.0 beta – the new features of tomorrow’s Spark, today

Another thing I’ve been itching to announce is our new Spark 4 beta image. This new beta image joins our collection of Spark 3 images – and whilst the beta image isn’t eligible for official support from Canonical, it gives you an easy way to try out the latest upstream Apache Spark 4 beta features today!

Some of the new features of Spark 4 include:

Preview today using Charmed Spark and our Spark container image

You can freely access our Apache Spark 4 beta container image in Github Container Registry right here –

https://github.com/canonical/charmed-spark-rock/pkgs/container/charmed-spark/280005099?tag=4.0-22.04_edge

If you’d like to learn more about getting enterprise-grade support for Apache Spark from Canonical, contact us and we’ll be happy to jump on a call with you to discuss further, or you can browse our Charmed Spark product page if you prefer.

Ubuntu Server Admin

Next Why is Ubuntu Linux the leading choice to replace CentOS for financial services? »

Previous « 20 years of partnership: how our partners help us take Ubuntu across industries, markets and devices

Ubuntu-Mate

Ubuntu MATE 25.04 Release Notes

Ubuntu MATE 25.04 is ready to soar! 🪽 Celebrating our 10th anniversary as an official…

4 days ago

Ubuntu Weekly Newsletter Issue 887

Welcome to the Ubuntu Weekly Newsletter, Issue 887 for the week of April 6 –…

5 days ago

Apache Spark 4.0 beta release – try it now

A data warehouse, on a cloud-native data lake with Apache Kyuubi

Spark 4.0 beta – the new features of tomorrow’s Spark, today

Preview today using Charmed Spark and our Spark container image

Recent Posts

Canonical Releases Ubuntu 25.04 Plucky Puffin

Ubuntu 25.04 (Plucky Puffin) Released

Extended Security Maintenance for Ubuntu 20.04 (Focal Fossa) begins May 29, 2025

Ubuntu 20.04 LTS End Of Life – activate ESM to keep your fleet of devices secure and operational

Ubuntu MATE 25.04 Release Notes

Ubuntu Weekly Newsletter Issue 887

Apache Spark 4.0 beta release – try it now

A data warehouse, on a cloud-native data lake with Apache Kyuubi

Spark 4.0 beta – the new features of tomorrow’s Spark, today

Preview today using Charmed Spark and our Spark container image

Related Post

Recent Posts

Canonical Releases Ubuntu 25.04 Plucky Puffin

Ubuntu 25.04 (Plucky Puffin) Released

Extended Security Maintenance for Ubuntu 20.04 (Focal Fossa) begins May 29, 2025

Ubuntu 20.04 LTS End Of Life – activate ESM to keep your fleet of devices secure and operational

Ubuntu MATE 25.04 Release Notes

Ubuntu Weekly Newsletter Issue 887

This Website Uses Cookies