#

gRPC Programming Tutorials & Engineering Articles

175 gRPC tutorials, guides, and engineering insights from NVIDIA, Uber, Netflix, and more

gRPC Articles & Tutorials

Filter:
Google logo
Google
Intermediate
Real-time AI agents break traditional request-response load balancing paradigms because they rely on long-lived, stateful bidirectional streams that obscure true server capacity. To solve this, developers must implement application-level session tracking directly within the runtime to accurately measure the committed concurrent workload of active conversations. By feeding these precise session counts alongside standard CPU utilization metrics into a hybrid routing algorithm, infrastructure can effectively distribute stateful AI traffic and prevent individual backend bottlenecks.
Simerus Mahesh
8 min read
Includes Code
--
ClickHouse logo
ClickHouse
Intermediate
A walkthrough of adding OpenTelemetry instrumentation to two ASP.NET services — an Order API and a Payment Service — and shipping traces, logs, and metrics to ClickStack, with auto-correlated signals and cross-service trace waterfalls out of the box.
NVIDIA logo
NVIDIA
Advanced
The path from a trained AI model to production should be smooth, but rarely is. Many teams invest weeks fine-tuning models, only to discover that exporting to a…
Lovina Dmello
10 min read
Includes Code
--
Google logo
Google
Intermediate
Google Cloud has introduced a high-performance integration that connects Rapid Storage directly to PyTorch via the fsspec interface to eliminate AI training bottlenecks. By utilizing Google’s Colossus architecture and bidirectional gRPC streaming, the solution offers up to 15 TiB/s aggregate throughput and significant reductions in latency. These improvements allow developers to speed up total training time by 23% with zero code changes required beyond updating the storage bucket type.
Trinadh Kotturu, Martin Durant
4 min read
Includes Code
--
OpenAI logo
OpenAI
Intermediate
By Brian Yu and Ashwin Nathan, Members of the Technical Staff
OpenAI Team
8 min read
Includes Code
--
Stripe logo
Stripe
Intermediate
In this post we’ll dive into how we built a virtual cashless payment method that works for prepaid and Stripe-issued credits, and cleanly integrates with our accounting, compliance, and other frameworks.
Pratik Gupta
11 min read
Includes Code
--
Shopify logo
Shopify
Advanced
SimGym infrastructure for synthetic customers at scale. A Shopify + NVIDIA collaboration.
Javier Moreno
9 min read
Includes Code
--
Spotify logo
Spotify
Advanced
The article discusses Spotify's multi-agent architecture designed to enhance advertising workflows by addressing structural issues within their ad business.
Pratik Rasam and Ralph Sylvain
10 min read
Has Summary
--
Cloudflare logo
Cloudflare
Intermediate
The article discusses the ecdysis library developed by Cloudflare, which enables graceful restarts for Rust services without dropping live connections.
Manuel Olguín Muñoz
10 min read
Includes Code
Has Summary
--
Netflix Technology Blog
15 min read
Includes Code
--
Uber logo
Uber
Intermediate
This article introduces uForwarder, Uber's open-source push-based consumer proxy for Apache Kafka's async queuing system.
Zhifeng Chen, Yang Yang, Haifeng Chen
12 min read
Has Summary
--
LinkedIn logo
LinkedIn
Advanced
The article discusses how LinkedIn developed the Contextual Agent Playbooks & Tools (CAPT) to enhance AI coding agents with organizational context, enabling them to better assist engineers in their...
Ajay Prakash
17 min read
Has Summary
--
NVIDIA logo
NVIDIA
Intermediate
The article discusses NVIDIA's Alpamayo, a comprehensive ecosystem designed for developing reasoning-based autonomous vehicle (AV) systems.
Marco Pavone
11 min read
Includes Code
Has Summary
--
NVIDIA logo
NVIDIA
Advanced
The article discusses how to simulate an accurate radio environment for 5G and 6G systems using the NVIDIA Aerial Omniverse Digital Twin (AODT).
Tommaso Balercia
10 min read
Includes Code
Has Summary
--
Uber logo
Uber
Advanced
This article discusses how Uber utilizes a pull-based ingestion model in OpenSearch™ to effectively index streaming data.
Yupeng Fu, Varun Bharadwaj, Shuyi Zhang, Xu Xiong, Michael Froh
14 min read
Has Summary
--
NVIDIA logo
NVIDIA
Advanced
The article discusses the evolution of AI data centers into AI factories and the necessity for advanced telemetry solutions like NVIDIA Spectrum-X Ethernet to optimize AI workloads.
Berkin Kartal
7 min read
Has Summary
--
NVIDIA logo
NVIDIA
Advanced
The article discusses the transformation of AI-native 6G network design through the NVIDIA Aerial Omniverse Digital Twin, emphasizing the need for a dynamic, continuous integration approach to Radi...
Tommaso Balercia
7 min read
Has Summary
--
LinkedIn logo
LinkedIn
Advanced
The article discusses how LinkedIn enhanced its recommendation systems using SGLang, an open-source LLM serving framework.
Steven Shimizu
10 min read
Includes Code
Has Summary
--
OpenAI Team
7 min read
Includes Code
--
NVIDIA logo
NVIDIA
Intermediate
The article discusses the role of the distributed User Plane Function (dUPF) in the evolution of telecommunications towards 6G, emphasizing its importance for enabling ultra-low latency and high th...
Yuyong Zhang
9 min read
Has Summary
--
Stripe logo
Stripe
Intermediate
This article discusses how to connect Stripe events to frontend applications using various integration methods, enhancing user experience through real-time updates.
ClickHouse logo
ClickHouse
Intermediate
ClickHouse version 25. 8 introduces 45 new features, 47 performance optimizations, and 119 bug fixes, enhancing its capabilities as a high-performance analytical database.
ClickHouse Team
15 min read
Includes Code
Has Summary
--
ClickHouse logo
ClickHouse
Advanced
The article explores the potential of Large Language Models (LLMs) to replace Site Reliability Engineers (SREs) in performing root cause analysis (RCA) for production issues.
Lionel Palacin and Al Brown
76 min read
Includes Code
Has Summary
--
Uber logo
Uber
Advanced
The article discusses the evolution of Uber's Search Platform, highlighting its transition from Elasticsearch to an in-house solution called Sia, and ultimately to the adoption of OpenSearch.
Yupeng Fu, Shubham Gupta, Shanshan Song, Mingmin Chen
15 min read
Has Summary
--
Google logo
Google
Intermediate
The article announces the general availability of Gemini Code Assist in Apigee API Management, highlighting its AI-assisted capabilities for API development.
Sujin Park, Roderick Griner
3 min read
Has Summary
--
Netflix logo
Netflix
Advanced
The article discusses Netflix's Unified Data Architecture (UDA), which aims to streamline data modeling across its various platforms by allowing teams to define business concepts once and use them ...
Netflix Technology Blog
18 min read
Has Summary
--
Google logo
Google
Beginner
Google Cloud has announced the general availability of the Apigee APIM Operator, which enhances API management capabilities within Google Kubernetes Engine (GKE).
Sanjay Pujare
2 min read
Includes Code
Has Summary
--
NVIDIA logo
NVIDIA
Advanced
The article discusses advancements in Federated Learning (FL) specifically in the context of large language models (LLMs), focusing on the challenges of communication overhead and memory constraint...
Ziyue Xu
8 min read
Has Summary
--
Google logo
Google
Advanced
The article announces the general availability of the Apigee Extension Processor, a new capability that enhances Apigee's ability to manage and secure backend services and modern application archit...
Ishita Saxena, Sanjay Pujare
9 min read
Has Summary
--
NVIDIA logo
NVIDIA
Beginner
The article discusses the integration of physical AI and autonomous systems in industrial operations through the use of digital twins.
Ashley Goldstein
5 min read
Has Summary
--
NVIDIA logo
NVIDIA
Intermediate
The article discusses the integration of Flower and NVIDIA FLARE, two significant frameworks in the federated learning ecosystem.
Holger Roth
8 min read
Includes Code
Has Summary
--
NVIDIA logo
NVIDIA
Advanced
The article discusses NVIDIA DGX Cloud Serverless Inference, an auto-scaling AI inference solution that simplifies the deployment and scaling of AI applications across multi-cloud and on-premises e...
Vishal Ganeriwala
9 min read
Has Summary
--
SafetyCulture logo
SafetyCulture
Intermediate
SafetyCulture describes how they streamlined gRPC development for their offline-first mobile apps by automating code generation for their C++ middleware layer called Crux.
Chan Ryu
5 min read
Includes Code
Has Summary
--
ClickHouse logo
ClickHouse
Advanced
This article discusses the open sourcing of kubenetmon, a tool developed by ClickHouse to monitor data transfer in ClickHouse Cloud.
Ilya Andreev
24 min read
Includes Code
Has Summary
--
This article details SafetyCulture's comprehensive approach to secure string input validation in microservices, covering the four essential steps: decode, normalize/canonicalize, sanitize, and vali...
Peter Arts
23 min read
Includes Code
Has Summary
--
LinkedIn logo
LinkedIn
Intermediate
The article discusses LinkedIn's implementation of a Stateful Workload Operator for managing stateful systems on Kubernetes.
Michael Youssef
14 min read
Has Summary
--
NVIDIA logo
NVIDIA
Advanced
The article discusses the introduction of AI-RAN technology by NVIDIA, which aims to revolutionize telecom infrastructure by integrating AI capabilities into Radio Access Networks (RAN).
Soma Velayutham
13 min read
Has Summary
--
NVIDIA logo
NVIDIA
Intermediate
The article discusses the release of new Unreal Engine 5 on-device plugins for NVIDIA ACE, aimed at simplifying and scaling AI-powered MetaHuman character deployment on Windows PCs.
Ike Nnoli
4 min read
Has Summary
--