Articles tagged: real-time

38 articles

Local models

NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics

NVIDIA's Cosmos-H-Dreams enables real-time generative simulation for surgical robotics, allowing local models to train on synthetic, high-...

Jul 29, 20266 min
AI agents

Agentic AI vs Generative AI: The Shift from Static Outputs to Autonomous Action

Agentic AI systems, built on generative models, move beyond content creation to autonomous decision-making. This article explores how agen...

Jul 24, 20268 min
Local models

The State of Simulation for Physical AI: An Overview

Simulation is transforming physical AI by enabling safe, scalable training for robots and autonomous systems. Local models bring real-time...

Jul 22, 20267 min
AI agents

NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval

NVIDIA's Nemotron-3-8B-Embedding model achieves the top ranking on the Retrieval Text Embedding Benchmark (RTEB), setting a new standard f...

Jul 17, 20269 min
Local models

Mistral Updates: Enhanced Local AI Models for Edge Deployment

Mistral AI unveils new versions of its open-weight models optimized for local execution, featuring improved efficiency, lower latency, and...

Jul 16, 20266 min
AI tools

Introducing Real World VoiceEQ: Measuring the Human Quality of Voice AI

Real World VoiceEQ is a new benchmark that evaluates voice AI systems on human-likeness, emotional expressiveness, and natural prosody, pr...

Jul 16, 20267 min
Local models

Mistral’s Latest Updates: New Local Models and Enhanced Efficiency

Mistral AI has unveiled new local models optimized for on-device inference, offering improved speed and lower memory usage. These updates...

Jul 13, 20267 min
AI agents

Agentic AI vs Generative AI: The Shift from Content Creation to Autonomous Action

Generative AI creates content; Agentic AI acts on it. This article explores how combining GPT models with autonomous agents enables dynami...

Jul 12, 20267 min
AI tools

That Is Embarrassing: Why Frontier AI Still Makes Things Up, and What to Do About It

Frontier AI models continue to generate plausible-sounding but false information, a persistent flaw known as hallucination. This article e...

Jul 12, 20267 min
AI agents

From Generative AI to Agentic AI: The Next Frontier in Autonomous Systems

Generative AI creates content; Agentic AI takes action. This article explores how combining large language models with autonomous decision...

Jul 11, 20266 min
Local models

Mistral's Latest Updates: New Local Models and Performance Gains

Mistral AI releases enhanced local models with improved efficiency, reduced memory usage, and better reasoning. Discover key updates, benc...

Jul 9, 20267 min
Local models

Mistral's Latest Updates: New Local Models and Enhanced Performance

Mistral AI unveils new local models with improved efficiency and accuracy, including Mistral 7B v2 and specialized variants. These updates...

Jul 8, 20267 min
AI tools

Introducing DeepLearning.AI Pro: Your Gateway to Advanced AI Mastery

DeepLearning.AI Pro unlocks premium courses, projects, and expert mentorship for AI professionals. From LLM specialization to MLOps, this...

Jul 8, 20267 min
AI agents

Remote Agents in Vibe: Powered by Mistral Medium 3.5

Discover how remote agents in Vibe, powered by Mistral Medium 3.5, enable decentralized, autonomous task execution. This article explores...

Jul 7, 20267 min
AI tools

Time-Series LLMs, Explained with t0-alpha

Time-series LLMs like t0-alpha leverage transformer architectures to analyze sequential data. This article explains how t0-alpha handles f...

Jul 5, 20266 min
AI agents

AI Agents Explained: What Is a ReAct Loop and How Does It Work?

The ReAct loop combines reasoning and acting to enable AI agents to solve complex tasks iteratively. By alternating between thought, actio...

Jul 4, 20268 min
Local models

Introducing Mistral OCR 4: A New Era for Local Text Recognition

Mistral OCR 4 brings state-of-the-art optical character recognition capabilities to local environments, offering high accuracy, fast infer...

Jul 3, 20268 min
Local models

Hugging Face and Cerebras bring Gemma 4 to real-time voice AI

Hugging Face and Cerebras collaborate to run Gemma 4 models for real-time voice AI on local hardware, enabling low-latency speech processi...

Jul 2, 20267 min
AI agents

Introducing Dynamic Subagents in Deep Agents

Dynamic subagents enhance AI agent systems by enabling real-time delegation of specialized tasks. This modular approach improves scalabili...

Jun 30, 20267 min
AI research

The Next Frontier in AI Research: Beyond Generative Models

AI research is shifting from scaling generative models to building efficient, reasoning-driven systems. New paradigms like neuro-symbolic...

Jun 26, 20268 min
AI agents

From Local LLM to Tool-Using Agent

Learn how to transform a local large language model into a powerful agent by integrating external tools like web search, APIs, and code ex...

Jun 26, 20268 min
Local models

Introducing Mistral OCR 4: Next-Gen Local OCR for AI Workflows

Mistral OCR 4 brings high-accuracy text extraction to local AI models, enabling offline document processing with superior layout detection...

Jun 25, 20268 min
AI research

The Frontier of Artificial Intelligence: Current Trends in AI Research

AI research is rapidly evolving, from large language models and multimodal systems to breakthroughs in reasoning and safety. This article...

Jun 21, 20267 min
AI agents

Tool Calling, Explained: How AI Agents Decide What to Do Next

Discover how AI agents use tool calling to decide their next action. This article breaks down the decision-making process, from function s...

Jun 21, 20269 min
Guides

Testing Mythos and Fable: Moving Beyond SWE-bench with Nvidia’s Open Contender

Explore how Nvidia’s new open-source framework challenges SWE-bench dominance. Learn to test AI models with Mythos and Fable for real-worl...

Jun 20, 20268 min
AI research

The Frontier of AI Research: Current Trends and Future Directions

AI research is advancing rapidly, from deep learning breakthroughs to foundational models. This article explores key trends like multimoda...

Jun 20, 20267 min
Guides

Building a Custom GStreamer Plugin for NVIDIA DeepStream

Learn how to build a custom GStreamer plugin for NVIDIA DeepStream. This guide covers the plugin structure, element registration, and prac...

Jun 19, 20267 min
AI research

MosaicLeaks: Can your research agent keep a secret?

MosaicLeaks reveals how AI research agents can inadvertently reconstruct sensitive information from fragmented data. This article explores...

Jun 19, 20268 min
AI research

Is it agentic enough? Benchmarking open models on your own tooling

Learn how to evaluate open-source AI agents for autonomy and task completion using custom benchmarks. A practical guide for researchers an...

Jun 18, 20269 min
AI research

The Frontiers of Artificial Intelligence: Current Trends in AI Research

AI research is rapidly evolving, focusing on areas like generative models, reinforcement learning, and ethical frameworks. These advances...

Jun 18, 20268 min
AI agents

Agentic Resource Discovery: Let Agents Search

Agentic Resource Discovery empowers AI agents to autonomously search, evaluate, and retrieve resources like APIs, datasets, or tools. This...

Jun 18, 20267 min
Guides

LLMs Inside the Product: A Practical Field Guide

A hands-on guide to integrating large language models into products, covering architecture patterns, prompt engineering, cost optimization...

Jun 17, 20266 min
Guides

Drilling Into AI’s Financial Sustainability

Explore the hidden costs of AI development and deployment, from hardware to energy. Learn practical strategies for budgeting, optimizing m...

Jun 17, 20266 min
AI coding

How We Made Coding Agent Spend Predictable

Discover how we engineered a cost-predictable coding agent by combining token budgets, early stopping, and adaptive context management. Le...

Jun 16, 20266 min
AI agents

GPU Time-Slicing for Concurrent LLM Agents on Kubernetes

Learn how GPU time-slicing enables concurrent LLM agents on Kubernetes, maximizing GPU utilization and reducing costs. This article covers...

Jun 14, 20266 min
AI research

olmo-eval: An evaluation workbench for the model development loop

olmo-eval is an evaluation workbench designed to integrate seamlessly into the model development loop, enabling rapid iteration and system...

Jun 12, 20267 min
AI tools

When GPU Utilization Lies: The Hidden Systems Problem Slowing Modern AI

A clear and practical article about artificial intelligence for a professional audience.

Jun 12, 20266 min
AI agents

How Benchling builds agents when the smartest AI isn't smart enough

A clear and practical article about artificial intelligence for a professional audience.

Jun 12, 20267 min