Articles tagged: local LLM

17 articles

Local models

GPU Management: Why Idle GPUs Are the New Grounded Aircraft

In local AI deployments, idle GPUs mirror grounded aircraft: they consume capital, occupy space, and depreciate without yielding returns....

Jul 31, 202611 min
Local models

Mistral AI’s New Local Models: Enhanced Performance and Broader Accessibility

Mistral AI releases updated local models with improved efficiency, lower memory usage, and better reasoning. The updates include new quant...

Jul 30, 20269 min
Local models

Mistral’s Latest Open-Source Models: Powering Local AI

Mistral releases new lightweight models optimized for on-device inference. Their latest updates improve performance, efficiency, and acces...

Jul 25, 20268 min
Local models

Mistral Boosts Local AI Performance with New Model Optimizations

Mistral's latest updates focus on local model deployment: new quantization methods reduce memory footprint for Mixtral 8x7B by 30%, and of...

Jul 21, 20269 min
Local models

Mistral's Latest Updates: Pushing Local AI Forward

Mistral has released new versions of its open-weight models, improving performance on local hardware. Updates include enhanced reasoning,...

Jul 18, 20266 min
AI agents

Agentic AI vs. Generative AI: Redefining Intelligent Systems

Generative AI creates content, while Agentic AI takes action. This article explores how combining these technologies enables autonomous ag...

Jul 14, 20267 min
Local models

Mistral’s Latest Updates: New Local Models and Enhanced Efficiency

Mistral AI has unveiled new local models optimized for on-device inference, offering improved speed and lower memory usage. These updates...

Jul 13, 20267 min
Local models

Mistral's Latest Updates: New Local Models and Open-Source Advances

Mistral AI has released new local models with improved efficiency and performance. These updates include enhanced reasoning capabilities a...

Jul 5, 20266 min
AI agents

AI Agents Explained: What Is a ReAct Loop and How Does It Work?

The ReAct loop combines reasoning and acting to enable AI agents to solve complex tasks iteratively. By alternating between thought, actio...

Jul 4, 20268 min
AI agents

From Local LLM to Tool-Using Agent

Learn how to transform a local large language model into a powerful agent by integrating external tools like web search, APIs, and code ex...

Jun 26, 20268 min
Local models

Introducing Mistral OCR 4: Revolutionizing Local Document Understanding

Mistral OCR 4 brings powerful optical character recognition to local models, enabling fast, private, and accurate text extraction from ima...

Jun 26, 20266 min
AI tools

An LLM as Arbiter in RAG Retrieval: Picking the Right Candidate with Reasons

Explore how to use an LLM as an intelligent arbiter to select the best document from RAG retrieval candidates, enhancing accuracy with con...

Jun 26, 20268 min
Local models

Introducing Mistral OCR 4: Next-Gen Local OCR for AI Workflows

Mistral OCR 4 brings high-accuracy text extraction to local AI models, enabling offline document processing with superior layout detection...

Jun 25, 20268 min
Local models

We Got Local Models to Triage the OpenClaw Repo for FREE!*

Discover how we used local AI models to automate issue triage on the OpenClaw repository at zero cost, enhancing efficiency and reducing m...

Jun 23, 20266 min
AI tools

Shipping huggingface_hub every week with AI, open tools, and a human in the loop

Discover how the huggingface_hub library is released weekly using AI for code review and open tools for automation, while keeping a human...

Jun 23, 20267 min
Local models

Run a Local LLM with OpenClaw on Your Mac Mini

Learn how to install and run OpenClaw on a Mac Mini for private, offline AI inference. Step-by-step guide covers setup, model loading, and...

Jun 16, 20266 min
AI agents

GPU Time-Slicing for Concurrent LLM Agents on Kubernetes

Learn how GPU time-slicing enables concurrent LLM agents on Kubernetes, maximizing GPU utilization and reducing costs. This article covers...

Jun 14, 20266 min