#Mixture-of-Experts

18 articles with this tag

Mistral says Large 4 leads Europe, proof thin
Artificial Intelligence

Mistral says Large 4 leads Europe, proof thin

Arthur Mensch pitched Mistral Large 4 on France Inter as an hours-long agentic model, but benchmarks, scores and broad access remain undisclosed.

about 2 hours ago
MoE Models Tackle LLM Hallucinations
AI Research

MoE Models Tackle LLM Hallucinations

InnerExpert leverages MoE architecture's internal signals for per-token hallucination detection, achieving state-of-the-art results with high efficiency.

about 2 months ago
llamafile v0.10.5 Ships With Big Local Models
Artificial Intelligence

llamafile v0.10.5 Ships With Big Local Models

Llamafile v0.10.5 adds support for two large local AI models, Ternary Bonsai 27B and Laguna-S-2.1, by updating its core llama.cpp integration.

2 months ago
Thinking Machines Lab cuts costs with Inkling-Small
Artificial Intelligence

Thinking Machines Lab cuts costs with Inkling-Small

Thinking Machines Lab launches Inkling-Small, a 276B parameter model that delivers comparable performance to its larger predecessor at a fraction of the cost.

2 months ago
Together AI partners with Moonshot AI
Technology

Together AI partners with Moonshot AI

Together AI partners with Moonshot AI to offer Kimi K3 and future models, providing developers with day-zero access to large-scale open-source AI.

2 months ago
Mira Murati Ships Inkling: 975B-Parameter Open Model Backed by Nvidia
AI Figures

Mira Murati Ships Inkling: 975B-Parameter Open Model Backed by Nvidia

Thinking Machines Lab released Inkling on July 15, a 975-billion-parameter open-weight mixture-of-experts model trained from scratch on 45 trillion tokens, 22 months after Mira Murati left OpenAI.

3 months ago
Together AI adds Inkling multimodal model
Technology

Together AI adds Inkling multimodal model

Together AI integrates Inkling, a new multimodal AI model from Thinking Machines Lab, offering text, image, and audio processing with controllable reasoning.

3 months ago
Inkling AI Model: Open-Weights Multimodality
AI Research

Inkling AI Model: Open-Weights Multimodality

Thinking Machines unveils Inkling, an open-weights, multimodal AI model with 975B parameters, designed for customization and efficient, controllable thinking.

3 months ago
Mira Murati's Interaction Model: Full-Duplex AI at 0.4 Seconds
AI Figures

Mira Murati's Interaction Model: Full-Duplex AI at 0.4 Seconds

Thinking Machines Lab's TML-Interaction-Small processes audio and video in 200ms chunks, responds in 0.4 seconds, and runs full-duplex without a VAD harness. Here is how the architecture works and what Murati said at Bloomberg Tech.

4 months ago
MobileMoE LLMs Redefine On-Device AI
AI Research

MobileMoE LLMs Redefine On-Device AI

MobileMoE LLMs redefine on-device AI, setting new performance and efficiency benchmarks for sub-billion parameter models on smartphones.

4 months ago
AI at Graduations & Claude's Blackmail Tactics
Artificial Intelligence

AI at Graduations & Claude's Blackmail Tactics

IBM experts discuss AI's evolving role, from college graduations to ethical dilemmas like LLM data corruption and potential 'blackmail' scenarios.

5 months ago
Shodh-MoE: Unlocking Universal SciML
AI Research

Shodh-MoE: Unlocking Universal SciML

Shodh-MoE's sparse activation architecture resolves multi-physics interference in SciML, enabling universal foundation models with guaranteed physical properties.

5 months ago
Google DeepMind Unveils Gemma 4 AI Models
Artificial Intelligence

Google DeepMind Unveils Gemma 4 AI Models

Google DeepMind releases Gemma 4, a new family of open-source AI models featuring advanced architectures, multimodal capabilities, and improved performance.

6 months ago
AI in Science: Faster Discovery, New Insights
AI Research

AI in Science: Faster Discovery, New Insights

AI is revolutionizing scientific research, from data analysis to hypothesis generation. Experts discuss how AI tools like LLMs are accelerating discovery while highlighting the continued importance of human expertise.

7 months ago
Bayesian Uncertainty for Foundation Models
AI Research

Bayesian Uncertainty for Foundation Models

Variational Mixture-of-Experts Routing (VMoER) offers a scalable Bayesian approach to uncertainty quantification in foundation models, achieving significant improvements with minimal computational overhead.

7 months ago
Arcee Trinity Large Breaks Cover
AI Research

Arcee Trinity Large Breaks Cover

Arcee.ai unveils Trinity Large, a 400B-parameter Mixture-of-Experts model engineered for inference efficiency and enterprise long-context use, alongside smaller variants.

8 months ago
GPT-OSS-Puzzle-88B: Faster AI, Same Brains
AI Research

GPT-OSS-Puzzle-88B: Faster AI, Same Brains

GPT-OSS-Puzzle-88B offers substantial inference speedups for large language models without sacrificing accuracy, utilizing techniques like MoE pruning and window attention.

8 months ago
Step 3.5 Flash: AI's New Efficiency Standard
AI Research

Step 3.5 Flash: AI's New Efficiency Standard

Step 3.5 Flash AI model revolutionizes AI efficiency with a 196B parameter foundation and 11B active parameters, offering competitive performance with lower latency.

8 months ago