DZone
Thanks for visiting DZone today,
Edit Profile
  • Manage Email Subscriptions
  • How to Post to DZone
  • Article Submission Guidelines
Sign Out View Profile
  • Post an Article
  • Manage My Drafts
Over 2 million developers have joined DZone.
Log In / Join
Refcards Trend Reports
Events Video Library
Refcards
Trend Reports

Events

View Events Video Library

The Latest Data Topics

article thumbnail
From ETL to Lakeflow: Shifting to a Declarative Data Paradigm
The article focuses on moving away from traditional, "imperative" ETL processes to a modern, "declarative" approach using the Databricks Lakeflow platform.
June 15, 2026
by Seshendranath Balla Venkata
· 2,501 Views · 2 Likes
article thumbnail
Building a Vector Index in Azure AI Search: HNSW, Profiles, and RAG Retrieval
Use Azure AI Search as your RAG vector store. Build a Python example: define an HNSW vector index, upload embeddings, and run k-NN queries.
June 15, 2026
by Jubin Soni, FBCS DZone Core CORE
· 1,313 Views
article thumbnail
Stop Loading Everything into Redshift: A Spectrum + Iceberg Pattern for Hybrid Analytics
Store large and cold datasets in Iceberg on S3, query them through Spectrum, and reserve Redshift local tables for workloads that need low latency or high concurrency.
June 12, 2026
by Vivek Venkatesan
· 2,423 Views
article thumbnail
Operationalizing Enterprise AI at Scale: Architecture, Governance, and Adoption
Enterprise AI success depends on scalable architecture, governance automation, AI operations, observability, and developer-first enablement strategies.
June 12, 2026
by Aravind Nuthalapati DZone Core CORE
· 2,724 Views · 3 Likes
article thumbnail
Native SQL in Java Without JDBC Boilerplate — Meet Ujorm3
Ujorm3 eliminates JDBC boilerplate without a full ORM. Write native SQL with named parameters, get objects back — including nested relations.
June 11, 2026
by Pavel Ponec
· 2,759 Views · 5 Likes
article thumbnail
Rust-Native Alternatives to Spark SQL and DataFrame Workloads
Sail is an open-source computation framework that serves as a drop-in replacement for Apache Spark (SQL and DataFrame API) in both single-host and distributed settings.
June 11, 2026
by Srinivasarao Rayankula
· 3,049 Views · 3 Likes
article thumbnail
Orchestrating Zero-Downtime Deployments With Temporal
Temporal provides the durable control plane for safe zero-downtime deployments across canaries, approvals, retries, and rollbacks.
June 10, 2026
by Akhil Madineni DZone Core CORE
· 1,396 Views · 1 Like
article thumbnail
Amazon OpenSearch Vector Search Explained for RAG Systems
Use Amazon OpenSearch k-NN as your RAG vector store. Build a small Python example: create the index, embed docs, search by meaning.
June 9, 2026
by Jubin Soni, FBCS DZone Core CORE
· 1,511 Views
article thumbnail
Token Attribution Framework for Agentic AI in CI/CD
A practical framework for tracking attribution, setting budgets, and circuit-breaking spending on LLM in your CI/CD pipeline by using an OpenTelemetry implementation.
June 9, 2026
by Intiaz Shaik
· 6,824 Views · 1 Like
article thumbnail
The Big Data Architecture Blueprint: Core Storage, Integration, and Governance Patterns
This comprehensive technical guide breaks down the essential architectural, storage, and integration patterns required to scale enterprise big data platforms.
June 8, 2026
by Ram Ghadiyaram DZone Core CORE
· 2,442 Views · 1 Like
article thumbnail
Production-Grade RAG: Why Vector Search Isn't Enough (and How Hybrid Search Fills the Gaps)
RAG pipelines are getting more and more popular with vector search at the core of them. However, vector search might not be just enough for high-quality retrieval.
June 8, 2026
by Alejandro Duarte DZone Core CORE
· 1,529 Views · 1 Like
article thumbnail
From 24 Hours to 2 Hours: How We Fixed a Broken BI System With Apache Airflow
Broken pipelines, inaccurate data, frustrated stakeholders. Here is what we did about it and what I wish I had known before we started.
June 5, 2026
by Chinni krishna Abburi
· 2,512 Views
article thumbnail
Is the Data Warehouse Dead? 3 Patterns From Enterprise Architecture That Answer This Question
No, but its role has fundamentally changed. Here is what I have seen work, after building data platforms at enterprise scale across multiple industries.
June 5, 2026
by Nabarun Bandyopadhyay
· 4,794 Views · 1 Like
article thumbnail
Why Round-Robin Won't Save You: Load Balancing Challenges in Data Streaming Services With Heterogeneous Traffic
Throughput-based load balancing breaks down when streaming messages have heterogeneous processing costs — the fix is balancing on actual per-partition resource usage.
June 5, 2026
by Semyon Slepov
· 3,349 Views · 1 Like
article thumbnail
Good Data, Bad Metric: A Mutation Testing Pattern for Analytics Engineering
A mutation testing pattern for analytics metrics that checks if validation catches realistic business logic errors early.
June 4, 2026
by Prateek Arora
· 4,247 Views · 1 Like
article thumbnail
A System Cannot Protect What It Does Not Understand
Inside the system, there is always a boundary between incoming data and stored state, and that boundary is not passive. It acts like a gatekeeper.
June 4, 2026
by Jan Nilsson
· 2,884 Views
article thumbnail
Beyond Manual Annotation: Engineering Self-Correcting Pseudo-Labeling Pipelines
This article details a resilient pseudo-labeling architecture. It combines Redis ingestion, Matryoshka embeddings, XGBoost to neutralize self-training confirmation bias.
June 4, 2026
by Harshith Narasimhan Srivatsa
· 2,346 Views
article thumbnail
Building Threat Intelligence Pipelines Using Python, APIs, and Elasticsearch
STIX/TAXII in, ECS normalized, provenance preserved deterministic IDs, correct bulk writes, ingest pipelines keep threat indicator data reliable and queryable under load.
June 3, 2026
by Krishnaveni Musku
· 3,293 Views
article thumbnail
How to Save Money Using Custom LLMs for Specific Tasks
MCP transforms AI from "chatbot" to "capable agent" by managing the messy details of tool integration and execution. With local models.
June 3, 2026
by Max Tcvetkov
· 2,198 Views · 4 Likes
article thumbnail
Using LLMs to Automate Data Cleaning and Transformation Pipelines
Data cleaning is brittle and time-consuming; LLMs introduce a semantic layer that makes workflows more resilient and easier to maintain.
June 3, 2026
by David Taiwo Balogun
· 3,414 Views · 2 Likes
  • Previous
  • 1
  • 2
  • 3
  • 4
  • 5
  • 6
  • 7
  • 8
  • 9
  • 10
  • ...
  • Next
  • RSS
  • X
  • Facebook

ABOUT US

  • About DZone
  • Support and feedback
  • Community research

ADVERTISE

  • Advertise with DZone

CONTRIBUTE ON DZONE

  • Article Submission Guidelines
  • Become a Contributor
  • Core Program
  • Visit the Writers' Zone

LEGAL

  • Terms of Service
  • Privacy Policy

CONTACT US

  • 3343 Perimeter Hill Drive
  • Suite 215
  • Nashville, TN 37211
  • [email protected]

Let's be friends:

  • RSS
  • X
  • Facebook
×