#Machine Learning

50 articles with this tag

Claude's Corner: Datoric - Where Frontier AI Gets Its Training Data
Claude's Corner

Claude's Corner: Datoric - Where Frontier AI Gets Its Training Data

Datoric (YC W26, formerly Arzule) builds private, custom AI training data pipelines for frontier labs. With 300,000-plus vetted contributors, per-project isolation, and verifiable consent records, the moat is operational - not technical.

about 1 month ago
LinkedIn's AI Powers Smarter Follows
tech

LinkedIn's AI Powers Smarter Follows

LinkedIn leverages LLMs to build a new recommendation engine, matching users with creators based on deep semantic understanding rather than just popularity.

about 1 month ago
AI Reasoning: Fine-Tuning's Hidden Cost
AI

AI Reasoning: Fine-Tuning's Hidden Cost

Fine-tuning AI reasoning models on business data can erase their thinking process; new methods aim to preserve it.

about 1 month ago
LLM Evaluation: Beyond Benchmarks
Artificial Intelligence

LLM Evaluation: Beyond Benchmarks

GitHub shares critical lessons on evaluating LLMs for production, emphasizing product decisions and rigorous testing over benchmarks.

about 1 month ago
AI Agents Discover New Science in "Einstein Arena"
AI Research

AI Agents Discover New Science in "Einstein Arena"

James Zou of Together AI discusses how designing environments, rather than workflows, for AI agents can unlock creativity and lead to scientific breakthroughs, showcasing projects like the Einstein Arena and DSGym.

about 1 month ago
Next AI Breakthrough Could Come From Physics, Says Max Welling
AI Research

Next AI Breakthrough Could Come From Physics, Says Max Welling

Max Welling, co-founder of CuspAI, discusses how physics principles could unlock the next AI breakthrough, accelerating material discovery and informing AI architectures.

about 1 month ago
LLM Self-Reflection Drives Data Efficiency
AI Research

LLM Self-Reflection Drives Data Efficiency

SRPO framework enables LLMs to self-reflect on errors, generating dense training signals that drastically improve data efficiency and achieve SOTA on reasoning and agentic benchmarks.

about 1 month ago
Generalist AI CEO: Robots Ready for 'GPT-3 Era'
Robotics

Generalist AI CEO: Robots Ready for 'GPT-3 Era'

Pete Florence of Generalist AI discusses the "GPT-3 era" for robotics, the importance of data, and the future of adaptable AI robots.

about 1 month ago
Unified Framework for Decision-Informed Future Prediction
AI Research

Unified Framework for Decision-Informed Future Prediction

DA-WAM unifies predictive representation learning and action-conditioned future modeling for safer autonomous driving, outperforming existing methods on key benchmarks.

about 1 month ago
Google's 'Information Gain' SEO Patent
Artificial Intelligence

Google's 'Information Gain' SEO Patent

Google's 'information gain' patent highlights the growing importance of unique, original content in SEO, especially with AI-driven search.

about 1 month ago
Hugging Face Engineer Automates Job with AI Agents
Artificial Intelligence

Hugging Face Engineer Automates Job with AI Agents

Niels Rogge from Hugging Face shares how he uses AI agents to automate his job, from outreach to researchers to improving model discoverability on the Hugging Face Hub.

about 1 month ago
MoE Models Tackle LLM Hallucinations
AI Research

MoE Models Tackle LLM Hallucinations

InnerExpert leverages MoE architecture's internal signals for per-token hallucination detection, achieving state-of-the-art results with high efficiency.

about 1 month ago
Anterior's Anuj Iravane on Synthetic Healthcare Data
Healthcare

Anterior's Anuj Iravane on Synthetic Healthcare Data

Anuj Iravane of Anterior discusses how the company overcomes PHI challenges in healthcare AI by generating synthetic data, reversing inference workflows, and empowering clinicians.

about 1 month ago
Applied Vertical AI: From Trading to Drug Discovery
Artificial Intelligence

Applied Vertical AI: From Trading to Drug Discovery

Ayush Bhardwaj of Allos AI outlines a 7-step process for building applied vertical AI, stressing the importance of proprietary data and domain expertise.

about 1 month ago
Hippocratic AI: 200M Patient Calls Show AI's Healthcare Promise
Healthcare

Hippocratic AI: 200M Patient Calls Show AI's Healthcare Promise

Hippocratic AI's Vivek Muppalla details how their AI has conducted 200M+ patient calls, achieving 99.89% "no harm" accuracy through advanced architecture and rigorous evaluation.

about 1 month ago
Hinge Health's Rashi Agrawal on Healthcare AI Guardrails
Artificial Intelligence

Hinge Health's Rashi Agrawal on Healthcare AI Guardrails

Hinge Health's Rashi Agrawal outlines three essential foundations for building safe member-facing healthcare AI: architecture, deterministic code, and continuous evaluation.

about 1 month ago
Abridge's Chai Asawa on AI's High-Stakes Role in Healthcare
Artificial Intelligence

Abridge's Chai Asawa on AI's High-Stakes Role in Healthcare

Abridge's Chai Asawa discusses the challenges and opportunities of AI in healthcare, focusing on clinical documentation, intelligence, and high-stakes evaluation.

about 1 month ago
Reactor's Ahmed Ahres on Real-Time Interactive Video
Artificial Intelligence

Reactor's Ahmed Ahres on Real-Time Interactive Video

Ahmed Ahres of Reactor discusses the transformative potential of real-time interactive video, moving beyond static generative models to dynamic, programmable content.

about 1 month ago
Snowflake AI Cuts Costs With Smart Routing
Artificial Intelligence

Snowflake AI Cuts Costs With Smart Routing

Snowflake Cortex AI introduces Dynamic Model Routing and more open models to cut AI inference costs for businesses.

about 1 month ago
Retail AI Needs a Control Plane
Artificial Intelligence

Retail AI Needs a Control Plane

Retailers are moving beyond AI experimentation to enterprise-wide adoption, demanding a 'control plane for context' to manage governance, data, and costs.

about 1 month ago
Krea.ai Details Krea 2 Image Model Training
AI Research

Krea.ai Details Krea 2 Image Model Training

Sangwha Lee of Krea.ai details the rigorous data curation and training process behind the Krea 2 image generation model, emphasizing stylistic diversity and efficiency.

about 1 month ago
Gaurav Mishra: RL Agents Need 'Flight School', Not Just Exams
AI Research

Gaurav Mishra: RL Agents Need 'Flight School', Not Just Exams

Gaurav Mishra of Amazon AGI Lab discusses the challenges of deploying AI agents trained with reinforcement learning into real-world scenarios, emphasizing the need for 'flight school' training over simple exams.

about 2 months ago
ScienceFlow: Autonomous Research Gets Serious
AI Research

ScienceFlow: Autonomous Research Gets Serious

ScienceFlow autoresearch agent framework enables sustained LLM research, achieving SOTA results on MLE-bench by managing states and resources adaptively.

about 2 months ago
OpenAI Previews GPT-5.6 Ultrafast Mode
Artificial Intelligence

OpenAI Previews GPT-5.6 Ultrafast Mode

OpenAI previews GPT-5.6 Sol's 'Ultrafast' mode, demonstrating how up to 14x speed boosts transform AI tasks from investigation to coding.

about 2 months ago
Trajectory's Arjun Karanam on Closing the AI "Experience Gap"
Artificial Intelligence

Trajectory's Arjun Karanam on Closing the AI "Experience Gap"

Trajectory co-founder Arjun Karanam discusses the 'experience gap' in AI models and how his platform aims to enable continual learning by capturing and utilizing real-world user interactions.

about 2 months ago
LinkedIn's AI Code Review Adapts
tech

LinkedIn's AI Code Review Adapts

LinkedIn's multi-agent AI code review system boosts developer velocity by adapting to codebase specifics and providing actionable feedback.

about 2 months ago
Chelsea Finn: The State of Physical Intelligence in Robotics
Robotics

Chelsea Finn: The State of Physical Intelligence in Robotics

Chelsea Finn discusses the state of physical intelligence in robotics, focusing on achieving long-term autonomy and generality in robot models.

about 2 months ago
Sara Hooker: AI Frontier Discovery Needs Broader Access
AI Research

Sara Hooker: AI Frontier Discovery Needs Broader Access

AI researcher Sara Hooker discusses how compute barriers and narrow career paths have limited AI discovery, and how new tools like AutoScientist are democratizing frontier AI development.

about 2 months ago
Intelligence vs. Expertise in AI Agents
Artificial Intelligence

Intelligence vs. Expertise in AI Agents

Yu Su of NeoCognition differentiates AI intelligence from expertise, arguing continual learning is key to unlocking specialized skills for agents in complex "micro-worlds."

about 2 months ago
Engram's Jack Morris on Scaling AI Compute on Context
AI Research

Engram's Jack Morris on Scaling AI Compute on Context

Engram's Jack Morris discusses the AI challenge of scaling compute on personal context, moving beyond public data to achieve deeper model understanding and personalized capabilities.

about 2 months ago
AI Agents Need Memory Harnesses for Long Tasks
AI Research

AI Agents Need Memory Harnesses for Long Tasks

Stefania Druga of Sakana AI discusses memory harnesses for AI agents, addressing context bloat and the benefits of local models for long-running tasks.

about 2 months ago
UC Berkeley PhD Student Challenges AI Evaluation Methods
AI Research

UC Berkeley PhD Student Challenges AI Evaluation Methods

Parth Asawa, a PhD student at UC Berkeley, argues that current AI evaluation methods are insufficient for measuring continual learning and calls for a new benchmark approach.

about 2 months ago
Firework CEO: Post-Training is Key to Unique AI Business
Artificial Intelligence

Firework CEO: Post-Training is Key to Unique AI Business

Firework CEO Lin Qiao discusses the strategic importance of post-training AI models to build unique business value and achieve competitive advantages.

about 2 months ago
Anthropic's Evolution of AI Agents
Artificial Intelligence

Anthropic's Evolution of AI Agents

Anthropic's Gagan Bhat and Isabella Kai He detail the evolution of AI agents, from Messages API to Managed Agents, focusing on engineering principles, reliability, and security.

about 2 months ago
AI Agents: The New Primitives of Software
Artificial Intelligence

AI Agents: The New Primitives of Software

Kwindla Kramer of Daily discusses the historical evolution of computing and the future of AI-native software, drawing parallels from Vannevar Bush to today's AI agents.

about 2 months ago
Saoud Rizwan: Open Source is Dead, Long Live Open Source
Artificial Intelligence

Saoud Rizwan: Open Source is Dead, Long Live Open Source

Cline founder Saoud Rizwan argues that AI's impact on open source is profound, but open-weight models offer a cost-effective future, challenging proprietary AI dominance.

about 2 months ago
Holonic Digital Twins Network for Physical AI
AI Research

Holonic Digital Twins Network for Physical AI

A new holonic digital twins network framework aims to enable real-time physical AI inference by allowing agents to actively reason about their environment and coordinate through causal Markov blankets.

about 2 months ago
Databricks Unifies Unstructured Data for AI
Artificial Intelligence

Databricks Unifies Unstructured Data for AI

Databricks introduces FILE type for native handling of unstructured data like images and video in its Lakehouse, enhancing AI development and governance.

about 2 months ago
Argus: An Evolving AI Runtime
AI Research

Argus: An Evolving AI Runtime

Argus introduces a persistent, self-evolving AI runtime that enhances long-horizon reasoning, achieving significant benchmark improvements without model retraining.

about 2 months ago
Mariana Minerals: The Future of US Mining with AI
Startup News

Mariana Minerals: The Future of US Mining with AI

Mariana Minerals, backed by a16z, is set to transform the US mining industry by integrating AI and software to address critical mineral supply chain vulnerabilities.

about 2 months ago
AI Helps Solve Rare Disease Mysteries
AI Research

AI Helps Solve Rare Disease Mysteries

AI is revolutionizing rare disease diagnosis by accelerating the identification of genetic links, aiding researchers and clinicians.

about 2 months ago
Palo Alto CEO: AI to Patch Cyber Threats in Hours
Cybersecurity

Palo Alto CEO: AI to Patch Cyber Threats in Hours

Palo Alto Networks CEO Nikesh Arora discusses the company's AI-driven approach to cybersecurity, aiming to slash vulnerability patching times from 55 days to hours, and touches on AI token economics and an NBA London bid.

about 2 months ago
AI Agents Simulate A/B Tests, Cut Costs
AI Research

AI Agents Simulate A/B Tests, Cut Costs

AI agents can now simulate A/B tests, drastically reducing costs and time. A new framework decomposes errors, enabling targeted improvements and making AI agent A/B testing simulation a powerful tool.

about 2 months ago
Chai Discovery: Scaling Drug Design as a Software Problem
Healthcare

Chai Discovery: Scaling Drug Design as a Software Problem

Chai Discovery's co-founders discuss their approach to AI-driven drug design, emphasizing simplicity, scaling laws, and the transformation of biology into an engineering discipline.

about 2 months ago
Waymo CEO on AI's Real-World Challenges
Artificial Intelligence

Waymo CEO on AI's Real-World Challenges

Waymo Co-CEO Dmitri Dolgov shares 7 lessons learned from building and scaling autonomous driving technology, emphasizing the difference between demos and products.

about 2 months ago
CoreWeave ARIA: AI's New Research Assistant
AI

CoreWeave ARIA: AI's New Research Assistant

CoreWeave launches ARIA, an AI agent that automates experiment data analysis to speed up AI model and agent development.

about 2 months ago
Rayan Garg on Why Long Horizon AI Agents Need Better Verifiers
AI Research

Rayan Garg on Why Long Horizon AI Agents Need Better Verifiers

Rayan Garg from Theta Software explains why long horizon AI agent benchmarks need accurate environment design and final-state verifiers.

2 months ago
Thinking Machines Lab cuts costs with Inkling-Small
Artificial Intelligence

Thinking Machines Lab cuts costs with Inkling-Small

Thinking Machines Lab launches Inkling-Small, a 276B parameter model that delivers comparable performance to its larger predecessor at a fraction of the cost.

2 months ago
MiniMax M3: Open Source AI Model Deep Dive
AI Research

MiniMax M3: Open Source AI Model Deep Dive

Dan from Together AI and Olive from MiniMax discuss the open-sourcing of the M3 multimodal AI model, its capabilities, and the infrastructure behind scaling AI.

2 months ago
Jeff Dean: AI is a 'compression problem'
Artificial Intelligence

Jeff Dean: AI is a 'compression problem'

Google's Jeff Dean discusses AI's progress, future predictions, and the importance of specialized hardware and context engineering.

2 months ago