June 2026 Alibaba Cloud Big Data & AI Newsletter: Tech updates, new releases, market trends, and customer practices.
Brought to you by the Alibaba Cloud Big Data & AI Product Team.
Fully managed cloud-native big data platform for elastic Data+AI lakehouse solutions.
This article describes how to use the spark-submit command line interface (CLI) to submit a Spark job after EMR Serverless Spark is connected to ECS.
This article introduces a data processing workflow that integrates Realtime Compute for Apache Flink, EMR Serverless Spark, and Apache Paimon to enable real-time data ingestion.
This artile introduces the usability and maintainability of EMR Serverless Spark in stream processing.
This article is compiled from the first session of the EMR StarRocks online open class - EMR Serverless StarRocks3.
This article introduces the integration of Paimon and Spark, specifically focusing on query optimization.
This article compares the performance of Paimon and Hudi on Alibaba Cloud EMR and explores their respective roles in building quasi-real-time data warehouses.
This article explains how to monitor big data in EMR using Prometheus Service.
The latest entry of the Open-Source Folks Talk discusses the history of the first Apache Incubation Project on Alibaba Cloud.
In this article, we’ll explain how to run map-reduce jobs in the Alibaba Cloud EMR Cluster.
In this article, we'll introduce how to create an Alibaba Cloud EMR cluster step by step.
This article describes how to optimize the performance of the product features provided by the Enterprise Edition to help you efficiently access lake houses.
This article shares the best practices of InMobi based on the open-source big data service of Alibaba Cloud.
This article describes the solution of an open-source real-time data warehouse based on EMR OLAP.
A guide to configure integration between Alibaba Cloud EMR with Active Directory.
This article explains the four stages of lake house evolution within the Shanghai Shuhe Group.
Big Data is among the biggest IT trends of the last years. Maintaining a large infrastructure for analytics is a major challenge for Big Data.
This article is an overview of the best practices for big data processing in Spark taken from a lecture.