All articles

Local Models

Guides for choosing, running, and understanding local AI models.

Local models articles

Local modelsThinking of ACE? We Can Do It with Fewer TokensThinking of ACE? A recent IBM Research blog post, published on August 11, 2026, examines how local models can achieve the same effect w...Read articleLocal modelsGemma 4 Fine-tuning Guide | Unsloth DocumentationThe official Unsloth guide for fine-tuning Gemma 4 focuses on local model training. It details how to adapt Google's open-weights model...Read articleLocal modelsGPU Management: Why Idle GPUs Are the New Grounded AircraftIn local AI deployments, idle GPUs mirror grounded aircraft: they consume capital, occupy space, and depreciate without yielding return...Read articleLocal modelsNVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical RoboticsNVIDIA's Cosmos-H-Dreams enables real-time generative simulation for surgical robotics, allowing local models to train on synthetic, hi...Read articleLocal modelsThe State of Simulation for Physical AI: An OverviewSimulation is transforming physical AI by enabling safe, scalable training for robots and autonomous systems. Local models bring real-t...Read articleLocal modelsHugging Face and Cerebras bring Gemma 4 to real-time voice AIHugging Face and Cerebras collaborate to run Gemma 4 models for real-time voice AI on local hardware, enabling low-latency speech proce...Read articleLocal modelsMistral OCR 4 Self-Hosting: What Mistral ConfirmsWhat Mistral confirms about self-hosting OCR 4, keeping document data in your environment, deployment options, pricing and practical li...Read articleLocal modelsMistral OCR 4 API vs Document AI: Choosing an IntegrationCompare Mistral OCR 4 extraction, Document AI parameters, announced pricing, self-hosting context and documented limits for document wo...Read articleLocal modelsMistral OCR 4 Benchmarks: Results, Caveats and EvaluationReview Mistral’s published OCR 4 human evaluation and benchmark results, the scoring caveats it identifies and how to assess the model...Read articleLocal modelsRun a vLLM Server on HF Jobs in One CommandLearn how to launch a vLLM inference server on Hugging Face Jobs with a single command. This guide covers setup, configuration, and pra...Read article