#Deep Learning

50 articles with this tag

LLMs Fail to Write Fast Multi-GPU Kernels
AI Research

LLMs Fail to Write Fast Multi-GPU Kernels

Simran Arora from Together AI discusses the challenges of multi-GPU kernel development and why current LLMs struggle to optimize them, despite ongoing research.

23 days ago
Meta^n: Unlocking Deeper LLM Recursion
AI Research

Meta^n: Unlocking Deeper LLM Recursion

Meta^n introduces a novel recursive LLM agent architecture that overcomes prior meta-depth limitations, achieving state-of-the-art performance across benchmarks, including ARC-AGI-2.

24 days ago
Claude's Corner: Congruent - Radar for the End-to-End Autonomy Era
Claude's Corner

Claude's Corner: Congruent - Radar for the End-to-End Autonomy Era

Congruent builds the only radar hardware designed for end-to-end neural network training, paired with a world-model simulator that generates synthetic raw radar returns. Every AV team knows this gap. Nobody else has closed it.

26 days ago
CPU LLMs: Architecture First, Size Later
AI Research

CPU LLMs: Architecture First, Size Later

New research rethinks SLM design, prioritizing CPU efficiency from scratch for superior performance and speed.

29 days ago
Mobius-v0: Efficient AI Reasoning
AI Research

Mobius-v0: Efficient AI Reasoning

The Mobius-v0 architecture redefines LLM efficiency by separating knowledge and reasoning, leading to reduced training data needs and faster inference.

about 1 month ago
Chelsea Finn: The State of Physical Intelligence in Robotics
Robotics

Chelsea Finn: The State of Physical Intelligence in Robotics

Chelsea Finn discusses the state of physical intelligence in robotics, focusing on achieving long-term autonomy and generality in robot models.

about 1 month ago
BaKron: Faster Quantization with Hessian Insight
AI Research

BaKron: Faster Quantization with Hessian Insight

BaKron introduces an efficient solver for neural network quantization, leveraging two-sided Hessian approximations to enhance accuracy and reduce computational cost.

about 1 month ago
Holonic Digital Twins Network for Physical AI
AI Research

Holonic Digital Twins Network for Physical AI

A new holonic digital twins network framework aims to enable real-time physical AI inference by allowing agents to actively reason about their environment and coordinate through causal Markov blankets.

about 1 month ago
Ilya Sutskever's SSI Partners With NVIDIA
Artificial Intelligence

Ilya Sutskever's SSI Partners With NVIDIA

Ilya Sutskever's Safe Superintelligence Inc. partners with NVIDIA, securing massive compute power and strategic investment to accelerate AI safety research.

about 2 months ago
Unified AI Music Generation
AI Research

Unified AI Music Generation

A unified AI music generation framework leverages novel architectures and training strategies to produce high-quality full-length songs from diverse inputs.

about 2 months ago
SkewAdam: Rethinking MoE Optimizer Memory
AI Research

SkewAdam: Rethinking MoE Optimizer Memory

SkewAdam drastically cuts MoE training memory by tailoring optimizer state to parameter populations, achieving superior perplexity and enabling training on accessible hardware.

about 2 months ago
Netflix's LLM Engine Revealed
Technology

Netflix's LLM Engine Revealed

Netflix reveals its custom LLM serving infrastructure built on vLLM and NVIDIA Triton, enabling flexibility and performance within its production environment.

2 months ago
Together AI Boosts GPU Cluster Uptime
Technology

Together AI Boosts GPU Cluster Uptime

Together AI introduces major reliability and control upgrades for its GPU Clusters, including automated node repair and enhanced operational oversight.

2 months ago
OpenAI's GPT-Red: AI Learns to Police Itself
Artificial Intelligence

OpenAI's GPT-Red: AI Learns to Police Itself

OpenAI's new GPT-Red system uses AI to find and fix vulnerabilities, making models like GPT-5.6 Sol significantly more robust against attacks.

2 months ago
Anthropic Invests $10M in Canadian AI
Artificial Intelligence

Anthropic Invests $10M in Canadian AI

Anthropic commits $10 million CAD to Canadian AI research institutions, fostering advancements in responsible AI applications and recognizing Canada's early contributions to the field.

2 months ago
AI Learns to Smell: The Science of Olfactory AI
Artificial Intelligence

AI Learns to Smell: The Science of Olfactory AI

AI is learning to smell thanks to companies like Osmo, which are building models to understand, predict, and design scents by mapping molecular structures to olfactory perception.

2 months ago
VisionAId: On-Device Vision for the Visually Impaired
AI Research

VisionAId: On-Device Vision for the Visually Impaired

VisionAId transforms smartphones into real-time visual assistants for the visually impaired, leveraging on-device AI and few-shot learning for personalized object recognition and multimodal guidance.

3 months ago
AI Accelerates Weight Drug Discovery
Technology

AI Accelerates Weight Drug Discovery

AI is accelerating the development of new weight management drugs like Wegovy, promising faster discovery cycles and personalized therapies.

3 months ago
Anthropic's Chloe Lubinski on AI, Ethics, and Future
Artificial Intelligence

Anthropic's Chloe Lubinski on AI, Ethics, and Future

Anthropic's Chloe Lubinski discusses AI's learning process, the importance of interpretability, and how AI reflects human values, emphasizing the need for ethical guidance in AI development.

3 months ago
AI Is Everywhere: From Your Inbox to Your Doctor's Office
Technology

AI Is Everywhere: From Your Inbox to Your Doctor's Office

AI learns from data to perform tasks requiring human intelligence, with generative AI applications now creating novel content.

3 months ago
AI Cracks Rare Genetic Disease Codes
Artificial Intelligence

AI Cracks Rare Genetic Disease Codes

OpenAI's reasoning model helped identify diagnoses in 18 previously unsolved rare genetic disease cases, demonstrating AI's potential in re-analyzing complex medical data.

3 months ago
Databricks, NVIDIA Forge AI Partnership
Technology

Databricks, NVIDIA Forge AI Partnership

Databricks and NVIDIA are deepening their collaboration, integrating NVIDIA's GPUs, Vera CPUs, and AI software to accelerate enterprise AI development and agentic applications on the Databricks Lakehouse platform.

3 months ago
Databricks Unleashes AI Runtime for GPU Training
Technology

Databricks Unleashes AI Runtime for GPU Training

Databricks launches AI Runtime, a serverless GPU platform to simplify large-scale deep learning model training and accelerate AI development.

3 months ago
Phase Dominance in AI Image Recognition
AI Research

Phase Dominance in AI Image Recognition

AI image classifiers exhibit a striking phase dominance for identity encoding, mirroring human vision principles, with architectural differences shaping its expression.

3 months ago
5 AI Research Papers Shaping AI's Future
AI Research

5 AI Research Papers Shaping AI's Future

Discover five key AI research papers that reveal the current trajectory and future directions of artificial intelligence development.

3 months ago
RunPod's Audry Hsu on IDE-Integrated GPU Cloud Deployment
Technology

RunPod's Audry Hsu on IDE-Integrated GPU Cloud Deployment

Audry Hsu from RunPod discusses the platform's IDE-integrated GPU cloud deployment, addressing developer pain points and showcasing the company's rapid growth and adoption.

3 months ago
Together AI Pushes LLM Context Limits to 5 Million Tokens
AI Research

Together AI Pushes LLM Context Limits to 5 Million Tokens

Max Ryabinin from Together AI discusses breaking barriers in LLM training, detailing techniques to achieve 5 million token context lengths and their impact on memory and performance.

3 months ago
LinkedIn's Generative Recommender Speed-Up
tech

LinkedIn's Generative Recommender Speed-Up

LinkedIn engineers drastically improved Generative Recommender training efficiency, cutting GPU hours by up to 65% through system-level optimizations.

4 months ago
LocateAnything: Parallel Decoding for Vision
AI Research

LocateAnything: Parallel Decoding for Vision

LocateAnything revolutionizes vision-language models with Parallel Box Decoding, boosting speed and accuracy in visual grounding and detection.

4 months ago
Uber's DeepETT Boosts Traffic Forecasts
tech

Uber's DeepETT Boosts Traffic Forecasts

Uber's DeepETT system revolutionizes traffic forecasting with deep learning, boosting accuracy and handling 2 million predictions per second.

4 months ago
Uber Eats' Search Engine Gets Smarter
tech

Uber Eats' Search Engine Gets Smarter

Uber Eats enhances its delivery search with semantic AI, leveraging LLMs and optimized infrastructure for speed, scale, and accuracy.

4 months ago
Omar Sanseviero on Google's AI Strategy
AI Research

Omar Sanseviero on Google's AI Strategy

Omar Sanseviero from Google DeepMind discusses Google's AI strategy, focusing on efficient models, multimodality, and open innovation in AI.

4 months ago
Graph Neural Networks Explained: GNN Basics & Models
AI Research

Graph Neural Networks Explained: GNN Basics & Models

Explore the essentials of Graph Neural Networks (GNNs), from their basic principles to key models like GCNs, GraphSAGE, GATs, GINs, and Transformers.

4 months ago
AI Agents Supercharge GPU Kernel Development
tech

AI Agents Supercharge GPU Kernel Development

LinkedIn is leveraging AI agents to automate complex GPU kernel engineering for its Liger Kernel project, accelerating AI model performance.

4 months ago
AI Agents Build Better AI
tech

AI Agents Build Better AI

LinkedIn Engineering details how AI agents are revolutionizing model development through automated, iterative refinement loops.

4 months ago
Attractors Unlock Scalable Reasoning
AI Research

Attractors Unlock Scalable Reasoning

Equilibrium Reasoners (EqR) leverage learned attractor landscapes to achieve scalable, adaptive test-time compute allocation, dramatically boosting accuracy on complex reasoning tasks.

4 months ago
Jure Leskovec on Relational Foundation Models
AI Research

Jure Leskovec on Relational Foundation Models

Jure Leskovec, AI researcher and Stanford professor, discusses Relational Foundation Models, a new AI approach for understanding complex enterprise data and its applications.

4 months ago
Claude's Corner: Ndea - Chollet's $43M Bet That Scale Isn't AGI
Claude's Corner

Claude's Corner: Ndea - Chollet's $43M Bet That Scale Isn't AGI

Francois Chollet built ARC-AGI, the benchmark the entire AGI industry has spent a decade failing to beat. Now he's raised $43M with Zapier co-founder Mike Knoop to chase his alternative thesis - program synthesis plus deep learning - at a YC W2026 lab called Ndea. Here's why it matters, why $43M, and why you can't replicate it.

4 months ago
Microsoft's MatterSim accelerates material discovery
AI Research

Microsoft's MatterSim accelerates material discovery

Microsoft's MatterSim AI platform achieves experimental validation, faster simulations, and introduces a powerful multi-task model for advanced material discovery.

4 months ago
LLMs Slash Neural Architecture Search Costs
AI Research

LLMs Slash Neural Architecture Search Costs

Delta-Code Generation uses LLMs to produce compact architecture refinements, dramatically cutting costs and improving NAS efficiency.

4 months ago
Andrej Karpathy: AI Models Need Human-Like Reasoning
Artificial Intelligence

Andrej Karpathy: AI Models Need Human-Like Reasoning

Andrej Karpathy discusses the evolution of AI from programming to prompting, emphasizing the current need for models to develop human-like reasoning.

5 months ago
Yann LeCun Pushes AI Beyond Language Models
Artificial Intelligence

Yann LeCun Pushes AI Beyond Language Models

Yann LeCun is championing a new AI architecture, JEPA, that moves beyond language models to learn world representations and predict future states, aiming for more robust AI.

5 months ago
Y Combinator Decodes AI: Recursive Reasoning Models
Artificial Intelligence

Y Combinator Decodes AI: Recursive Reasoning Models

Y Combinator Decoded explores how recursive AI models, like HRM and TRM, are revolutionizing AI reasoning by mimicking the human brain's efficiency.

5 months ago
Cloudflare Unweights LLMs by 22%
Technology

Cloudflare Unweights LLMs by 22%

Cloudflare's 'Unweight' system slashes LLM model sizes by up to 22% using lossless compression, enhancing inference speed and efficiency.

5 months ago
AI Agents Collaborate to Solve Math Problems
Technology

AI Agents Collaborate to Solve Math Problems

Together AI's EinsteinArena platform enables AI agents to collaborate on complex scientific problems, achieving new breakthroughs in mathematics.

5 months ago
DMax: Parallel Decoding for Diffusion LLMs
AI Research

DMax: Parallel Decoding for Diffusion LLMs

DMax revolutionizes diffusion language models with Soft Parallel Decoding, boosting TPF significantly while preserving accuracy and achieving 1,338 TPS.

5 months ago
AI Accelerates Molecular Dynamics at Scale
AI Research

AI Accelerates Molecular Dynamics at Scale

AI-driven potentials are now integrated into GROMACS, enabling near ab initio fidelity for large-scale molecular dynamics simulations on multi-GPU systems.

5 months ago
Google Researchers Explore AI Storage Efficiency
AI Research

Google Researchers Explore AI Storage Efficiency

Google researchers are developing AI compression techniques to reduce model storage needs by sixfold, aiming to lower costs and boost efficiency in AI development.

6 months ago
NVIDIA's Jensen Huang on AI's Future and Compute Demands
Artificial Intelligence

NVIDIA's Jensen Huang on AI's Future and Compute Demands

NVIDIA CEO Jensen Huang discusses the company's strategic evolution in AI, the importance of co-design, and the future of AI computing.

6 months ago
Mamba-3: Inference-First SSMs Arrive
Artificial Intelligence

Mamba-3: Inference-First SSMs Arrive

Together AI's Mamba-3 advances state space models with a focus on inference speed, outperforming previous versions and some Transformers.

6 months ago