#
Elasticsearch Programming Tutorials & Engineering Articles
151 Elasticsearch tutorials, guides, and engineering insights from Netflix, Uber, ClickHouse, and more
Companies Using This
Elasticsearch Articles & Tutorials
Filter:
Aravind Selvan, Calder Lund, and Gaurang Sadekar
13 min read
Includes Code
--
Learn how to send structured .NET logs directly to ClickHouse using Serilog — with full schema control, full-text search, and SQL queries over your log data.
23 min read
Includes Code
--
High-performance full-text search, even on object storage. This deep dive from the ClickHouse engineers explains the new text index design and how it keeps queries fast at scale.
37 min read
Includes Code
--
Ten best practices for getting the most out of ClickHouse, from primary key design and data types to materialized views, ReplacingMergeTree, and join optimization — illustrated with benchmarks on a 150M row dataset.
ClickHouse now brings full-text search and large-scale analytics together in one engine, making it a powerful alternative to Elasticsearch for log analytics. This benchmark shows why.
24 min read
Includes Code
--
A walkthrough of adding OpenTelemetry instrumentation to two ASP.NET services — an Order API and a Payment Service — and shipping traces, logs, and metrics to ClickStack, with auto-correlated signals and cross-service trace waterfalls out of the box.
14 min read
Includes Code
--
A video analytics AI agent that can perceive, reason, and act based on massive amounts of video footage must be integrated with existing workflows and…
Tanya Lenz
13 min read
Includes Code
--
7 min read
--
Palantir
16 min read
Includes Code
--
Netflix Technology Blog
18 min read
Includes Code
--
ElasticsearchengineeringGitLabGitLab CIGPTJSONJWTKubernetesNode.jsOAuthPrometheusReactRustTypeScriptYAML
21 min read
Includes Code
--
Justin Lee, Adam Hudson
10 min read
--
Netflix Technology Blog
11 min read
Includes Code
--
Here’s how we made the search experience better, faster, and more resilient for GHES customers.
David Tippett
6 min read
Includes Code
--
Uber Engineering details their migration from a legacy monolithic monitoring system to a modern, cloud-native observability platform for their corporate network infrastructure.
Razvan Cicu, Giovanni Pepe
9 min read
Has Summary
--
The article discusses the use of AI Model Distillation to create efficient financial data workflows, focusing on the optimization of large language models (LLMs) for applications in quantitative fi...
Dhruv Desai
10 min read
Includes Code
Has Summary
--
The article discusses Shopify's innovative approach to building a high-performance product search engine that integrates Machine Learning (ML) models with C++ speed.
Mikhail Shakhray
6 min read
Includes Code
Has Summary
--
The article discusses the creation of a website for tracking team activity across GitHub repositories, initially intended as a single report but evolved into a comprehensive tool for comparing vari...
The article discusses the potential of lakehouses using open table formats like Apache Iceberg and Delta Lake for observability, highlighting their advantages in scalability, cost-effectiveness, an...
Melvyn Peignon & Dale McDiarmid
24 min read
Includes Code
Has Summary
--
This article discusses how Netflix built a resilient data platform using a Write-Ahead Log (WAL) to address data consistency, reliability, and operational efficiency challenges at scale.
The article discusses how Palantir optimizes Elasticsearch to enhance its defensive capabilities against poor access patterns, particularly focusing on indexing refresh semantics.
Palantir
18 min read
Includes Code
Has Summary
--
The article discusses the rising costs associated with observability in software engineering and proposes a shift towards open, cost-efficient architectures.
Mike Shi
13 min read
Has Summary
--
The article discusses Uber's implementation of encryption at rest and disk isolation at scale using their Stateful Platform, Odin.
Ivan Shibitov, Johan Abildskov
14 min read
Has Summary
--
The article discusses the advancements in NVIDIA cuVS, a GPU-accelerated vector search library designed for high-performance indexing and low-latency retrieval.
Corey Nolet
7 min read
Has Summary
--
The article discusses the NVIDIA AI Blueprint for Building Data Flywheels, which aims to optimize AI agents powered by large language models by reducing inference costs and improving latency.
Sylendran Arunagiri
2 min read
Has Summary
--
The article discusses the evolution of Uber's Search Platform, highlighting its transition from Elasticsearch to an in-house solution called Sia, and ultimately to the adoption of OpenSearch.
Yupeng Fu, Shubham Gupta, Shanshan Song, Mingmin Chen
15 min read
Has Summary
--
The article discusses the evolution of ClickHouse's observability platform, LogHouse, as it scales beyond 100 petabytes of data.
Rory Crispin, Dale McDiarmid
30 min read
Includes Code
Has Summary
--
The article discusses the NVIDIA AI Blueprint for building efficient AI agents through model distillation, focusing on the challenges of scaling intelligent applications and managing inference cost...
Daniel Glogowski
10 min read
Includes Code
Has Summary
--
The article discusses how NVIDIA Air Services can connect simulations with real-world data center infrastructure, enhancing capabilities and performance.
Sophia Schuur
6 min read
Includes Code
Has Summary
--
This article details how GitHub rebuilt its Issues search system to support nested queries with boolean AND/OR operators and parentheses.
Deborah Digges
10 min read
Includes Code
Has Summary
--
The article discusses how ClickHouse efficiently queries Parquet files, a key storage format for Lakehouse architectures, without requiring data ingestion.
The article discusses the One Billion Documents JSON Challenge, comparing the performance of ClickHouse against other popular databases like MongoDB, Elasticsearch, DuckDB, and PostgreSQL in storin...
Tom Schreiber
33 min read
Includes Code
Has Summary
--
The article 'Break Stuff on Purpose' discusses the importance of intentionally causing failures in systems to improve recovery processes and enhance resilience.
Sean Madden
8 min read
Has Summary
--
The article discusses the evolution of SQL-based observability, focusing on ClickHouse's advancements over the past year.
Dale McDiarmid & Ryadh Dahimene
25 min read
Includes Code
Has Summary
--
Netflix's TimeSeries Data Abstraction Layer is designed to efficiently store and query vast amounts of temporal event data with low latency.
Netflix Technology Blog
22 min read
Includes Code
Has Summary
--
The article discusses how NVIDIA optimizes data center performance using AI agents and the OODA loop strategy.
Aaron Erickson
11 min read
Has Summary
--
This article discusses the modernization of Uber's logging infrastructure using CLP, focusing on the development of an end-to-end system for managing unstructured logs.
Gao Xin, Jack Luo, Kirk Rodrigues
16 min read
Has Summary
--
The article discusses NVIDIA Metropolis, a platform for real-time vision AI that streamlines deployment through microservices and workflows.
Monika Jhuria
11 min read
Has Summary
--
This article introduces the multi-camera tracking workflow developed by NVIDIA, aimed at optimizing processes in large spaces such as warehouses and airports.
Monika Jhuria
11 min read
Includes Code
Has Summary
--
This article compares ClickHouse and Elasticsearch in terms of performance for large-scale data analytics, particularly focusing on `count(*)` aggregations over billions of rows.
This article compares ClickHouse and Elasticsearch, focusing on their mechanics for count aggregations.
Tom Schreiber
17 min read
Includes Code
Has Summary
--
The article discusses the implementation of reverse search functionality within Netflix's Graph Search, which allows users to find queries that match specific documents instead of the traditional m...
Netflix Technology Blog
9 min read
Includes Code
Has Summary
--
The article discusses strategies to minimize on-call burnout through effective alert observability, emphasizing the importance of actionable alerts and the analysis of alert data.
Monika Singh
12 min read
Includes Code
Has Summary
--
This article discusses the management of ClickHouse schemas as code using the Atlas tool, highlighting the transition from schema-less technologies to structured data management.
Rotem Tamir
6 min read
Includes Code
Has Summary
--
This article discusses Uber's experience with garbage collection (GC) tuning to enhance the reliability of Presto, an open-source distributed SQL query engine.
Cristian Velazquez, Vineeth Karayil Sekharan
11 min read
Has Summary
--
The article discusses how Uber utilizes Apache Pinot for real-time analytics of mobile app crashes, enhancing their ability to detect and resolve issues quickly.
Kriti Dangi, Anil Purohit, Parijat Bansal, Rohit Yadav
17 min read
Has Summary
--
The article discusses the experiences of interns in Slack's Data Engineering team, highlighting their impactful projects such as the Reliable Data Discovery Tool and the Job Performance Tracking an...
Camryn McDonald
10 min read
Has Summary
--
The article discusses the NVIDIA DOCA GPUNetIO library, which enables real-time network processing by leveraging GPU parallelism to optimize packet acquisition and transmission.
Elena Agostini
13 min read
Has Summary
--
Cadence 1. 0 is a powerful open-source workflow orchestration platform designed for building and managing stateful services at scale.