#AI Efficiency
6 articles with this tag

OpenAI GPT-5.6: Smarter, Cheaper
OpenAI's new GPT-5.6 family cuts costs and boosts performance through significant inference and agentic harness optimizations.

Rachel Nabors: Local AI Models for Frontier Results
Rachel Nabors advocates for using smaller, on-device AI models, showcasing their efficiency, cost savings, and performance benefits over large frontier models.
Scaling Agent Collaboration via Recursion
RecursiveMAS scales agent collaboration via a unified latent-space recursive computation, achieving significant accuracy gains with improved efficiency.

Perplexity CTO on GPT-5.5 Efficiency
Perplexity CTO Denis Yarats reveals GPT-5.5's impressive efficiency, using 56% fewer tokens for complex tasks and enabling faster user feedback.
Agentic Models Bypass Tool Reliance
HDPO framework enables agentic multimodal models to drastically reduce tool use by decoupling accuracy and efficiency optimization, fostering self-reliance without performance loss.

Step 3.5 Flash: AI's New Efficiency Standard
Step 3.5 Flash AI model revolutionizes AI efficiency with a 196B parameter foundation and 11B active parameters, offering competitive performance with lower latency.