Install
AI Reasoning at Scale Try out DeepSeek-R1âs reasoning capabilities with NVIDIA-hosted APIs or deploy it anywhere with NVIDIA NIM inference microservices. Accelerate Apache Spark ML on NVIDIA GPUs with Zero Code Change How Using a Reranking Microservice Can Improve Accuracy and Costs of Information Retrieval Superchar
- 39articles · 30d
- 3+ day agolatest article
- Aug 17, 2026earliest in window
- 10%with images
- 245avg words
- Science & Technology 39
- Computers & Electronics 35
- Software Dev. 33
- Hardware 3
- Science & Nature 3
- Business & Industrial 1
- Finance 1
- Jobs & Education 1
Please confirm you are human
This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.
A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.
News
From Wafer-Out to First Token: Codifying Supply Chain Expertise with Nemotron and Palantir Foundry
4+ day, 3+ hour ago (1065+ words) With highly dynamic availability, NVIDIA must decide what and how much material to allocate to each manufacturing site. This is known as the critical material allocation problem, and it is manually reworked every week. The allocation runs through the current…...
Run NVIDIA BioNeMo NIM Microservices for Protein Structure Prediction in Claude Science
1+ week, 6+ day ago (767+ words) NVIDIA BioNeMo Agent Toolkit closes that gap. The toolkit packages more than a decade of NVIDIA BioNeMo life sciences models, libraries, and workflows into agent-callable skills for biology, chemistry, genomics, and drug discovery. Built to run with any agent framework,…...
NVIDIA BlueField-4 Powers New Scale-In Network Infrastructure for Agentic AI Factories
2+ week, 6+ day ago (205+ words) North-south networks provide the access path into and out of a data center, connecting users, applications, data sources, storage systems, and services to individual systems. Traditional cloud data centers built these networks around software-defined infrastructure, composability, and elasticity, so resources…...
Solving Agentic AI Fleet Challenges with NVIDIA Vera CPU
3+ week, 2+ day ago (289+ words) The shape of an agentic trajectory is defined by its length and width (Figure 2). This is why the relevant optimization target for an agentic CPU fleet is the total number of completed user sessions, not raw core count. High-core-count systems…...
NVIDIA Vera Rubin and Blackwell Set a New Standard for Agentic AI Performance per Watt
2+ week, 6+ day ago (88+ words) AI agents have expanded inference from single-turn interactions into multi-step workflows that reason, invoke tools, coordinate subagents…...
NVIDIA AVO Reaches 100% on ARC-AGI-3, Demonstrating a Frontier-Level General-Purpose Architecture for Long-Horizon Autonomous Agents
3+ week, 2+ day ago (651+ words) A frontier language model is only one component of an AI agent. The surrounding agent system—often called a harness—determines how the model receives context…...
How Generative Recommenders Are Redefining RecSys at Scale
3+ week, 5+ day ago (695+ words) In production, RecSys models are often served online to millions of users under strict service-level agreements (SLAs), where small increases in latency can impact user experience. Unlike LLM workloads that may tolerate autoregressive decoding latency, RecSys models must frequently retrieve…...
Evaluating AI Agent Skill Performance with NVIDIA SkillEvaluator
3+ week, 5+ day ago (607+ words) AI agents are only as effective as the context they receive. Even with capable models and well-documented NVIDIA libraries, agents can spend extra steps finding…...