#DeepSeek-R1
3 articles with this tag

AI
AI Reasoning: Fine-Tuning's Hidden Cost
Fine-tuning AI reasoning models on business data can erase their thinking process; new methods aim to preserve it.
about 1 month ago

Artificial Intelligence
OpenAI Unveils Jalapeño Chip
OpenAI reveals Jalapeño, its custom AI inference chip, demonstrating industry-leading performance and signaling a move towards greater infrastructure control.
about 2 months ago

AI Figures
Liang Wenfeng's $294K DeepSeek-R1 RL Breakthrough Reached Nature
How Liang Wenfeng's DeepSeek-R1 used Group Relative Policy Optimization and pure reinforcement learning to produce emergent reasoning capabilities for $294,000 in training compute, and why the paper reached the cover of Nature.
3 months ago