DEV Community

#bigdata

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Spring Batch: Processing Large Data Volumes Efficiently (2026-09-03 17:31)

Spring Batch: Processing Large Data Volumes Efficiently (2026-09-03 17:31)

Comments
2 min read
Apache Data Lakehouse Weekly: July 29 to August 5, 2026

Apache Data Lakehouse Weekly: July 29 to August 5, 2026

Comments
27 min read
Apache Spark & PySpark, Explained Without a PhD

Apache Spark & PySpark, Explained Without a PhD

7
Comments
7 min read
The pipeline passed. The data still needs proof.

The pipeline passed. The data still needs proof.

Comments 2
3 min read
Understanding the ORC File Format: Why Is It So Fast

Understanding the ORC File Format: Why Is It So Fast

3
Comments
8 min read
Managing Hadoop Without the Headache: Introducing Hadoop Handler

Managing Hadoop Without the Headache: Introducing Hadoop Handler

2
Comments
3 min read
[EN] Data Driven Series #1: Apache Kafka and Elastic Stack Big Data Use Cases

[EN] Data Driven Series #1: Apache Kafka and Elastic Stack Big Data Use Cases

2
Comments
4 min read
Apache Data Lakehouse Weekly: July 21 to July 29, 2026

Apache Data Lakehouse Weekly: July 21 to July 29, 2026

1
Comments
26 min read
Top 12 Spark Interview Problems for Data Engineers, With Answers

Top 12 Spark Interview Problems for Data Engineers, With Answers

Comments
10 min read
Does Early Dragon Control Help Scaling Compositions? A Big Data Analysis of League of Legends

Does Early Dragon Control Help Scaling Compositions? A Big Data Analysis of League of Legends

9
Comments 4
4 min read
Predictive Maintenance in 2026: How AI, Edge Computing, and Agentic Systems Turn Detection Into Action

Predictive Maintenance in 2026: How AI, Edge Computing, and Agentic Systems Turn Detection Into Action

Comments
14 min read
How Uber Built Its Big Data System — From a Few TBs to 350 Petabytes with Sub-Hour Latency

How Uber Built Its Big Data System — From a Few TBs to 350 Petabytes with Sub-Hour Latency

Comments
9 min read
Your warehouse isn't expensive. Your full table scans are.

Your warehouse isn't expensive. Your full table scans are.

1
Comments 1
5 min read
Migrating a ScyllaDB Cluster the “Brain Transplant” Way

Migrating a ScyllaDB Cluster the “Brain Transplant” Way

Comments
6 min read
"We Have DevOps, So Why Not DataOps?"

"We Have DevOps, So Why Not DataOps?"

Comments
2 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.