DZone
Thanks for visiting DZone today,
Edit Profile
  • Manage Email Subscriptions
  • How to Post to DZone
  • Article Submission Guidelines
Sign Out View Profile
  • Post an Article
  • Manage My Drafts
Over 2 million developers have joined DZone.
Log In / Join
Refcards Trend Reports
Events Video Library
Refcards
Trend Reports

Events

View Events Video Library

The Latest Data Topics

article thumbnail
Parallel Kafka Batch Processing With Kotlin Coroutines in Spring Boot
Learn how Kotlin Coroutines improve Spring Boot Kafka batch processing with parallel execution, resource throttling, and faster database operations.
June 16, 2026
by Erkin Karanlık
· 2,651 Views · 1 Like
article thumbnail
Cutting Data Pipeline Costs and Data Freshness Issues With Netflix Maestro and Apache Iceberg: A Practical Tutorial
Iceberg replaces filesystem state with a metadata tree (cheap queries, ACID snapshots). Maestro replaces cron with event signals (fresh data).
June 16, 2026
by Intiaz Shaik
· 5,406 Views · 3 Likes
article thumbnail
Reducing RAG Hallucinations With Relationship-Aware Retrieval
An architectural idea and how it addresses the retrieval weaknesses that lead to hallucinations, with a reference implementation using RudraDB.
June 16, 2026
by Mahesh Vaijainthymala Krishnamoorthy
· 1,471 Views
article thumbnail
From ETL to Lakeflow: Shifting to a Declarative Data Paradigm
The article focuses on moving away from traditional, "imperative" ETL processes to a modern, "declarative" approach using the Databricks Lakeflow platform.
June 15, 2026
by Seshendranath Balla Venkata
· 2,294 Views · 2 Likes
article thumbnail
Building a Vector Index in Azure AI Search: HNSW, Profiles, and RAG Retrieval
Use Azure AI Search as your RAG vector store. Build a Python example: define an HNSW vector index, upload embeddings, and run k-NN queries.
June 15, 2026
by Jubin Abhishek Soni DZone Core CORE
· 1,248 Views
article thumbnail
Stop Loading Everything into Redshift: A Spectrum + Iceberg Pattern for Hybrid Analytics
Store large and cold datasets in Iceberg on S3, query them through Spectrum, and reserve Redshift local tables for workloads that need low latency or high concurrency.
June 12, 2026
by Vivek Venkatesan
· 2,259 Views
article thumbnail
Operationalizing Enterprise AI at Scale: Architecture, Governance, and Adoption
Enterprise AI success depends on scalable architecture, governance automation, AI operations, observability, and developer-first enablement strategies.
June 12, 2026
by Aravind Nuthalapati DZone Core CORE
· 2,517 Views · 3 Likes
article thumbnail
Native SQL in Java Without JDBC Boilerplate — Meet Ujorm3
Ujorm3 eliminates JDBC boilerplate without a full ORM. Write native SQL with named parameters, get objects back — including nested relations.
June 11, 2026
by Pavel Ponec
· 2,695 Views · 5 Likes
article thumbnail
Rust-Native Alternatives to Spark SQL and DataFrame Workloads
Sail is an open-source computation framework that serves as a drop-in replacement for Apache Spark (SQL and DataFrame API) in both single-host and distributed settings.
June 11, 2026
by Srinivasarao Rayankula
· 2,836 Views · 3 Likes
article thumbnail
Orchestrating Zero-Downtime Deployments With Temporal
Temporal provides the durable control plane for safe zero-downtime deployments across canaries, approvals, retries, and rollbacks.
June 10, 2026
by Akhil Madineni
· 1,324 Views
article thumbnail
Amazon OpenSearch Vector Search Explained for RAG Systems
Use Amazon OpenSearch k-NN as your RAG vector store. Build a small Python example: create the index, embed docs, search by meaning.
June 9, 2026
by Jubin Abhishek Soni DZone Core CORE
· 1,264 Views
article thumbnail
Token Attribution Framework for Agentic AI in CI/CD
A practical framework for tracking attribution, setting budgets, and circuit-breaking spending on LLM in your CI/CD pipeline by using an OpenTelemetry implementation.
June 9, 2026
by Intiaz Shaik
· 6,635 Views · 1 Like
article thumbnail
The Big Data Architecture Blueprint: Core Storage, Integration, and Governance Patterns
This comprehensive technical guide breaks down the essential architectural, storage, and integration patterns required to scale enterprise big data platforms.
June 8, 2026
by Ram Ghadiyaram DZone Core CORE
· 2,189 Views · 1 Like
article thumbnail
Production-Grade RAG: Why Vector Search Isn't Enough (and How Hybrid Search Fills the Gaps)
RAG pipelines are getting more and more popular with vector search at the core of them. However, vector search might not be just enough for high-quality retrieval.
June 8, 2026
by Alejandro Duarte DZone Core CORE
· 1,330 Views · 1 Like
article thumbnail
From 24 Hours to 2 Hours: How We Fixed a Broken BI System With Apache Airflow
Broken pipelines, inaccurate data, frustrated stakeholders. Here is what we did about it and what I wish I had known before we started.
June 5, 2026
by Chinni krishna Abburi
· 2,424 Views
article thumbnail
Is the Data Warehouse Dead? 3 Patterns From Enterprise Architecture That Answer This Question
No, but its role has fundamentally changed. Here is what I have seen work, after building data platforms at enterprise scale across multiple industries.
June 5, 2026
by Nabarun Bandyopadhyay
· 4,559 Views · 1 Like
article thumbnail
Why Round-Robin Won't Save You: Load Balancing Challenges in Data Streaming Services With Heterogeneous Traffic
Throughput-based load balancing breaks down when streaming messages have heterogeneous processing costs — the fix is balancing on actual per-partition resource usage.
June 5, 2026
by Semyon Slepov
· 3,311 Views · 1 Like
article thumbnail
Good Data, Bad Metric: A Mutation Testing Pattern for Analytics Engineering
A mutation testing pattern for analytics metrics that checks if validation catches realistic business logic errors early.
June 4, 2026
by Prateek Arora
· 4,201 Views · 1 Like
article thumbnail
A System Cannot Protect What It Does Not Understand
Inside the system, there is always a boundary between incoming data and stored state, and that boundary is not passive. It acts like a gatekeeper.
June 4, 2026
by Jan Nilsson
· 2,852 Views
article thumbnail
Beyond Manual Annotation: Engineering Self-Correcting Pseudo-Labeling Pipelines
This article details a resilient pseudo-labeling architecture. It combines Redis ingestion, Matryoshka embeddings, XGBoost to neutralize self-training confirmation bias.
June 4, 2026
by Harshith Narasimhan Srivatsa
· 2,314 Views
  • Previous
  • 1
  • 2
  • 3
  • 4
  • 5
  • 6
  • 7
  • 8
  • 9
  • 10
  • ...
  • Next
  • RSS
  • X
  • Facebook

ABOUT US

  • About DZone
  • Support and feedback
  • Community research

ADVERTISE

  • Advertise with DZone

CONTRIBUTE ON DZONE

  • Article Submission Guidelines
  • Become a Contributor
  • Core Program
  • Visit the Writers' Zone

LEGAL

  • Terms of Service
  • Privacy Policy

CONTACT US

  • 3343 Perimeter Hill Drive
  • Suite 215
  • Nashville, TN 37211
  • [email protected]

Let's be friends:

  • RSS
  • X
  • Facebook
×