How NVIDIA Uses Claude
30 engineering articles about Claude from NVIDIA's engineering team
Other NVIDIA Technologies
Other Companies Using Claude
Articles
Filter:
Learn how NVIDIA NeMo Switchyard routes AI agent workloads across models using tuning-free and tunable routers that balance model capability, cost, and latency.
Michelle Horton
11 min read
--
NVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they…
Tanya Lenz
4 min read
--
Developers building video analytics applications across large spaces must track the same object as it moves between camera views. Single-camera 2D tracking…
Elizabeth Goodman
11 min read
Includes Code
--
What if autonomous coding AI agents could push your vision reasoning models above 90% accuracy with almost no manual effort? When adapting vision reasoning…
Tanya Lenz
11 min read
Includes Code
--
AI scientists are emerging as a new interface for scientific computing. These agents can read papers, write code, generate hypotheses, call APIs, inspect files…
Kyle Tretina
9 min read
Includes Code
--
Training a speech AI model to correctly recognize or synthesize clinical terminology is surprisingly difficult. Drug names like Acetaminophen, Amlodipine…
John Jahanipour
12 min read
Includes Code
--
Deploy Self-Evolving Agents for Faster, More Secure Research with a Hermes Agent and NVIDIA NemoClaw
AI agents are a powerful tool for synthesizing data to accelerate research, summarize information, and help teams make decisions faster.
Agent harnesses like Claude Code, Codex, and LangChain Deep Agents are excellent orchestrators. They manage sessions, chain tools, execute code…
Autonomous AI agents are becoming more capable. Open models, Model Context Protocol (MCP)-connected tools, and portable skills are also making agents easier to…
Moshe Abramovitch
7 min read
Includes Code
--
In today’s data-driven world, organizations increasingly rely on video to capture critical information, yet extracting meaningful, real-time insights from…
Samuel Ochoa
11 min read
Includes Code
--
An agentic exchange must preserve a structured interaction: assistant turns interleave reasoning with one or more tool calls, and subsequent user turns return…
Generative AI’s explosive first chapter was defined by humans sending requests and models responding. The agentic chapter is different. Agents don’t follow a…
Eduardo Alvarez
11 min read
--
NVIDIA CUDA Tile (cuTile) is a tile-based programming model that enables developers to write GPU kernels in terms of tile-level operations—loads, stores…
In March 2026, three LLM agents generated over 600,000 lines of code, ran 850 experiments, and helped secure a first-place finish in a Kaggle playground…
Chris Deotte
7 min read
Includes Code
--
Coding agents are starting to write production code at scale. Stripe’s agents generate 1,300+ PRs per week. Ramp attributes 30% of merged PRs to agents.
Ishan Dhanani
16 min read
Includes Code
--
Developing real-time vision AI applications presents a significant challenge for developers, often demanding intricate data pipelines, countless lines of code…
NVIDIA Ising is the world’s first family of open AI models for building quantum processors, launching with two model domains: Ising Calibration and Ising…
Tom Lubowe
9 min read
--
Physical AI—AI systems that perceive, reason, and act in physically grounded simulated environments—is changing how teams design and validate robots and…
AI has evolved from assistants following your directions to agents that act independently. Called claws, these agents can take a goal, figure out how to achieve…
In this post, we dive into one of the most critical workloads in modern AI: Flash Attention, where you’ll learn: Environment requirements: See the quickstart…
The article provides practical security guidance for sandboxing agentic workflows, emphasizing the importance of managing execution risk associated with AI coding agents.
The article discusses the development of small orchestration agents, specifically the ToolOrchestra method, which automates the selection and management of models and tools for task-solving in AI s...
The article discusses the benchmarking of AI coding assistants in writing efficient CUDA code using the ComputeEval framework.
Daniel Rodriguez
2 min read
Has Summary
--
The article discusses the dual role of AI-enabled developer tools, highlighting both their potential to accelerate coding and the security vulnerabilities they introduce.
The article discusses the transformative role of domain-adapted large language models (LLMs) with reasoning capabilities in accelerating battery research.
Rucha Apte
11 min read
Has Summary
--
The article discusses the benchmarking of agentic large language models (LLMs) and vision-language models (VLMs) using NVIDIA NIM and the BALROG benchmark suite.
Davide Paglieri
6 min read
Has Summary
--
ComputeEval is an open-source framework designed to evaluate Large Language Models (LLMs) on CUDA code generation, focusing on high-performance GPU programming.
Daniel Rodriguez
4 min read
Has Summary
--
The article discusses the development of a new reward model, Llama 3.
Zhilin Wang
3 min read
Has Summary
--
Writer has launched two domain-specific AI models, Palmyra-Med 70B and Palmyra-Fin 70B, enhancing NVIDIA NIM's capabilities in healthcare and finance.
The article features Lokman Abbas Turki, a researcher at Sorbonne University, who applies high performance computing (HPC) to complex mathematical finance problems and cryptography.
Brad Nemire
7 min read
Has Summary
--
You've reached the end! All 30 articles loaded.