Articles tagged: AI efficiency

2 articles

Local models

Thinking of ACE? We Can Do It with Fewer Tokens

Thinking of ACE? A recent IBM Research blog post, published on August 11, 2026, examines how local models can achieve the same effect with...

Aug 12, 202610 min
AI agents

Prompt Caching with Deep Agents

Prompt caching reduces latency and cost in AI agents by storing and reusing processed prompts. This technique enables faster multi-step re...

Jun 27, 20267 min