Articles tagged: NVIDIA

30 articles

AI tools

How Full-Stack NIM Optimizations Deliver 2.5x More Users on Nemotron 3 Ultra

NVIDIA's full-stack NIM optimizations on Nemotron 3 Ultra raise serving throughput enough to reach 2.5x more concurrent users per deployme...

Sep 11, 202611 min
AI tools

Backing 16 Green AI Projects in Asia-Pacific: A Regional Push for Sustainable Innovation

A verified primary source announces support for 16 green AI initiatives across Asia-Pacific, focusing on sustainability-driven machine lea...

Sep 7, 202613 min
AI tools

Mapping Global Methane Emissions from Space with Deep Learning

Google Research has published a detailed approach for mapping global methane emissions from space using deep learning. The method leverage...

Sep 2, 202612 min
Guides

NVIDIA NVLink Fusion Brings NVHBM to Next-Generation AI Infrastructure

NVIDIA NVLink Fusion introduces NVHBM to next-generation AI infrastructure, expanding high-bandwidth memory pooling and interconnect effic...

Aug 29, 202613 min
Guides

How to Choose Full-Stack Observability for NVIDIA AI Factories

Selecting the right full-stack observability solution for NVIDIA AI factories requires understanding GPU telemetry, cluster metrics, and a...

Aug 13, 202611 min
Local models

Gemma 4 Fine-tuning Guide | Unsloth Documentation

The official Unsloth guide for fine-tuning Gemma 4 focuses on local model training. It details how to adapt Google's open-weights models u...

Aug 8, 202612 min
AI agents

Deploy Local Agents Everywhere with LFM2.5-2.6B

Discover how LFM2.5-2.6B enables lightweight, privacy-preserving AI agents on edge devices. This compact model delivers strong reasoning a...

Aug 5, 202611 min
AI tools

The Python Ecosystem That Changed AI Development

From NumPy to PyTorch, Python's libraries and frameworks have transformed AI from research to production. This exploration reveals how a v...

Aug 2, 202610 min
Local models

GPU Management: Why Idle GPUs Are the New Grounded Aircraft

In local AI deployments, idle GPUs mirror grounded aircraft: they consume capital, occupy space, and depreciate without yielding returns....

Jul 31, 202611 min
Local models

NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics

NVIDIA's Cosmos-H-Dreams enables real-time generative simulation for surgical robotics, allowing local models to train on synthetic, high-...

Jul 29, 20266 min
Local models

The State of Simulation for Physical AI: An Overview

Simulation is transforming physical AI by enabling safe, scalable training for robots and autonomous systems. Local models bring real-time...

Jul 22, 20267 min
AI agents

NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval

NVIDIA's Nemotron-3-8B-Embedding model achieves the top ranking on the Retrieval Text Embedding Benchmark (RTEB), setting a new standard f...

Jul 17, 20269 min
AI tools

Behind the Scenes of Distributed Training and Why Your GPU Wiring Matters as Much as Your Strategy

Distributed training accelerates AI model development, but network topology and GPU interconnect often bottleneck performance. This articl...

Jul 11, 20267 min
AI agents

From Local LLM to Tool-Using Agent

Learn how to transform a local large language model into a powerful agent by integrating external tools like web search, APIs, and code ex...

Jun 26, 20268 min
AI agents

3 Agents. 3 LLMs. 1 Aging GPU: Engineering Parallel Inference on Bare Metal

Learn how to run three AI agents with separate LLMs simultaneously on a single outdated GPU. This article covers bare-metal parallel infer...

Jun 25, 20267 min
AI coding

Build Your Own Local AI Coding Agent with Gemma 4 and OpenCode

Learn to create a free, private AI coding agent on your own machine using Gemma 4 and OpenCode. This guide covers setup, configuration, an...

Jun 25, 20267 min
Local models

We Got Local Models to Triage the OpenClaw Repo for FREE!*

Discover how we used local AI models to automate issue triage on the OpenClaw repository at zero cost, enhancing efficiency and reducing m...

Jun 23, 20266 min
Guides

Testing Mythos and Fable: Moving Beyond SWE-bench with Nvidia’s Open Contender

Explore how Nvidia’s new open-source framework challenges SWE-bench dominance. Learn to test AI models with Mythos and Fable for real-worl...

Jun 20, 20268 min
Guides

Building a Custom GStreamer Plugin for NVIDIA DeepStream

Learn how to build a custom GStreamer plugin for NVIDIA DeepStream. This guide covers the plugin structure, element registration, and prac...

Jun 19, 20267 min
Guides

Proteins: A Mosaic Pattern to Rule Them All?

Explore how AI models like AlphaFold decode the mosaic patterns of proteins, revolutionizing drug discovery and bioengineering with practi...

Jun 18, 20267 min
Guides

LLMs Inside the Product: A Practical Field Guide

A hands-on guide to integrating large language models into products, covering architecture patterns, prompt engineering, cost optimization...

Jun 17, 20266 min
AI agents

From the Hugging Face Hub to Robot Hardware with Strands Agents and LeRobot

Discover how Strands Agents and LeRobot bridge the gap between AI models on Hugging Face Hub and real-world robot hardware, enabling seaml...

Jun 17, 20267 min
Guides

DeepSeek Sharpens Its Reasoning: DeepSeek-R1, an Affordable Rival to OpenAI’s o1

DeepSeek-R1 brings advanced reasoning capabilities at a fraction of the cost of OpenAI’s o1. Learn how this open-source model matches o1 i...

Jun 15, 20266 min
AI agents

GPU Time-Slicing for Concurrent LLM Agents on Kubernetes

Learn how GPU time-slicing enables concurrent LLM agents on Kubernetes, maximizing GPU utilization and reducing costs. This article covers...

Jun 14, 20266 min
Guides

How to Train a Scoring Model in the Age of Artificial Intelligence

Learn how to build a robust scoring model using AI, from data preparation to model evaluation. This guide covers key steps, practical exam...

Jun 13, 20267 min
AI tools

When GPU Utilization Lies: The Hidden Systems Problem Slowing Modern AI

A clear and practical article about artificial intelligence for a professional audience.

Jun 12, 20266 min
AI agents

Agentic AI / Generative AI

A clear and practical article about artificial intelligence for a professional audience.

Jun 8, 20266 min
AI tools

NVIDIA H100 on Replicate: what the May 2025 announcement confirms

A historical, sourced summary of Replicate’s May 2025 H100 announcement, including its stated availability and the limits of those time-bo...

Jun 7, 20264 min
Local models

Mistral AI partners with NVIDIA to accelerate open frontier models

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20264 min
AI tools

I Built a C++ Backend So My GPU Would Stop Eating Air

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20268 min