Articles tagged: Ollama

29 articles

Local models

Gemma 4 Fine-tuning Guide | Unsloth Documentation

The official Unsloth guide for fine-tuning Gemma 4 focuses on local model training. It details how to adapt Google's open-weights models u...

Aug 8, 202612 min
Local models

GPU Management: Why Idle GPUs Are the New Grounded Aircraft

In local AI deployments, idle GPUs mirror grounded aircraft: they consume capital, occupy space, and depreciate without yielding returns....

Jul 31, 202611 min
Local models

NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics

NVIDIA's Cosmos-H-Dreams enables real-time generative simulation for surgical robotics, allowing local models to train on synthetic, high-...

Jul 29, 20266 min
Local models

The State of Simulation for Physical AI: An Overview

Simulation is transforming physical AI by enabling safe, scalable training for robots and autonomous systems. Local models bring real-time...

Jul 22, 20267 min
AI tools

How I’m Making Sure My Analytics Career Doesn’t Get Eaten by AI

AI is transforming analytics, but instead of fearing it, I leverage it as a co-pilot. By focusing on strategic thinking, data storytelling...

Jul 21, 20268 min
AI tools

That Is Embarrassing: Why Frontier AI Still Makes Things Up, and What to Do About It

Frontier AI models continue to generate plausible-sounding but false information, a persistent flaw known as hallucination. This article e...

Jul 12, 20267 min
AI agents

AI Agents Explained: What Is a ReAct Loop and How Does It Work?

The ReAct loop combines reasoning and acting to enable AI agents to solve complex tasks iteratively. By alternating between thought, actio...

Jul 4, 20268 min
Local models

Mistral OCR 4 Self-Hosting: What Mistral Confirms

What Mistral confirms about self-hosting OCR 4, keeping document data in your environment, deployment options, pricing and practical limit...

Jun 29, 20268 min
Local models

Mistral OCR 4 API vs Document AI: Choosing an Integration

Compare Mistral OCR 4 extraction, Document AI parameters, announced pricing, self-hosting context and documented limits for document workf...

Jun 28, 20266 min
AI agents

From Local LLM to Tool-Using Agent

Learn how to transform a local large language model into a powerful agent by integrating external tools like web search, APIs, and code ex...

Jun 26, 20268 min
Local models

Run a vLLM Server on HF Jobs in One Command

Learn how to launch a vLLM inference server on Hugging Face Jobs with a single command. This guide covers setup, configuration, and practi...

Jun 26, 20266 min
AI tools

An LLM as Arbiter in RAG Retrieval: Picking the Right Candidate with Reasons

Explore how to use an LLM as an intelligent arbiter to select the best document from RAG retrieval candidates, enhancing accuracy with con...

Jun 26, 20268 min
AI agents

How To Give Your Agent Memory

Learn how to equip AI agents with memory using vector databases, conversation history, and structured storage. Practical techniques for pe...

Jun 25, 20267 min
Local models

We Got Local Models to Triage the OpenClaw Repo for FREE!*

Discover how we used local AI models to automate issue triage on the OpenClaw repository at zero cost, enhancing efficiency and reducing m...

Jun 23, 20266 min
AI tools

Shipping huggingface_hub every week with AI, open tools, and a human in the loop

Discover how the huggingface_hub library is released weekly using AI for code review and open tools for automation, while keeping a human...

Jun 23, 20267 min
AI agents

Agentic Resource Discovery: Let Agents Search

Agentic Resource Discovery empowers AI agents to autonomously search, evaluate, and retrieve resources like APIs, datasets, or tools. This...

Jun 18, 20267 min
Guides

Drilling Into AI’s Financial Sustainability

Explore the hidden costs of AI development and deployment, from hardware to energy. Learn practical strategies for budgeting, optimizing m...

Jun 17, 20266 min
Local models

Run a Local LLM with OpenClaw on Your Mac Mini

Learn how to install and run OpenClaw on a Mac Mini for private, offline AI inference. Step-by-step guide covers setup, model loading, and...

Jun 16, 20266 min
AI agents

GPU Time-Slicing for Concurrent LLM Agents on Kubernetes

Learn how GPU time-slicing enables concurrent LLM agents on Kubernetes, maximizing GPU utilization and reducing costs. This article covers...

Jun 14, 20266 min
Local models

We Should Train AI to Betray Its Users

A clear and practical article about artificial intelligence for a professional audience.

Jun 8, 20268 min
Local models

My AI Couldn’t See My Files

A clear and practical article about artificial intelligence for a professional audience.

Jun 7, 20269 min
Local models

fast.ai—Making neural nets uncool again

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20269 min
Local models

Is an Online Master’s Degree in AI a Good Idea?

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20264 min
Local models

Mistral AI partners with NVIDIA to accelerate open frontier models

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20264 min
Local models

Introducing Mistral Small 4

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20264 min
Local models

Emmi joins Mistral to accelerate the AI-native industry

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20264 min
Local models

Introducing physics AI at Mistral: the foundation for engineering acceleration.

A clear and practical article about artificial intelligence for a professional audience.

Jun 6, 20264 min
Local models

Latest updates from Mistral.

A clear and practical article about artificial intelligence for a professional audience.

Jun 5, 20264 min
Local models

Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI

A clear and practical article about artificial intelligence for a professional audience.

Jun 5, 20264 min