#AI Ethics

50 articles with this tag

Ufonia's AI: Shipping Healthcare Safely to a Million Patients
Artificial Intelligence

Ufonia's AI: Shipping Healthcare Safely to a Million Patients

Jared Joselowitz of Ufonia explains how to safely deploy healthcare AI to millions of patients using simulation, automated prompt optimization, and rigorous evaluation, bypassing traditional A/B testing.

6 days ago
AI Drone Swarms: The New Frontline of Warfare
Artificial Intelligence

AI Drone Swarms: The New Frontline of Warfare

Bloomberg reporter Katrin Manson discusses the rapid advancement of autonomous drone swarms in warfare, highlighting AI challenges and ethical concerns.

10 days ago
Sara Hooker: AI Frontier Discovery Needs Broader Access
AI Research

Sara Hooker: AI Frontier Discovery Needs Broader Access

AI researcher Sara Hooker discusses how compute barriers and narrow career paths have limited AI discovery, and how new tools like AutoScientist are democratizing frontier AI development.

13 days ago
Moonshot AI K3 Model Faces Distillation Claims
Startup News

Moonshot AI K3 Model Faces Distillation Claims

Moonshot AI's K3 model is accused of illicitly distilling Anthropic's Fable, sparking US concerns over intellectual property theft in AI.

about 1 month ago
Measuring True AI Autonomy
AI Research

Measuring True AI Autonomy

The Autonomous Agency Scale (AAS) moves beyond capability benchmarks to measure AI self-direction, revealing a significant 'idle gap' in current systems.

about 1 month ago
Kimi Verifier Rebuilds Trust in Open Source AI
AI

Kimi Verifier Rebuilds Trust in Open Source AI

Moonshot AI launches Kimi Vendor Verifier to ensure open-source AI models run accurately across all implementations, rebuilding trust in the ecosystem.

about 1 month ago
AI Vision Test Reveals Model Hallucinations
Artificial Intelligence

AI Vision Test Reveals Model Hallucinations

A new benchmark, PerceptionBench, reveals that even advanced multimodal AI models struggle with basic visual perception, often guessing answers instead of truly seeing.

about 1 month ago
Agentic AI Security: The New Math of Risk
Startup News

Agentic AI Security: The New Math of Risk

Agentic AI is redefining enterprise security, shifting from threat detection to a 'assume breach' recovery mindset, as AI-driven attacks outmatch traditional defenses.

about 2 months ago
Musk & Altman's AI Feud Reignites
Startup News

Musk & Altman's AI Feud Reignites

The Musk Altman feud 2026 intensified with an Apple lawsuit against OpenAI, reigniting public barbs and legal battles over AI's future.

about 2 months ago
AI Trust: Juries and Librarians as Solutions
Artificial Intelligence

AI Trust: Juries and Librarians as Solutions

Alex Bauer of Upside.tech proposes using 'juries' and 'librarians' to solve AI's trust problem, moving beyond simple hallucination fixes to a more robust approach.

about 2 months ago
OpenAI's GPT-Live Achieves Natural Conversation
Artificial Intelligence

OpenAI's GPT-Live Achieves Natural Conversation

OpenAI's technical staff unveil GPT-Live, showcasing its ability to engage in natural, nuanced conversations, moving beyond basic text generation.

about 2 months ago
Anthropic's Claude Mimics Human Brain Processing, Fuels AI Debate
AI Research

Anthropic's Claude Mimics Human Brain Processing, Fuels AI Debate

Anthropic's Claude AI model can mimic human brain processing, showing advanced reasoning and leading to discussions on AI consciousness and transparency.

about 2 months ago
UN Taps AI Leaders for New Global Commission
Technology

UN Taps AI Leaders for New Global Commission

Sakana AI's Ren Ito joins a new UN commission with tech and industry leaders to guide AI's global impact and trustworthiness.

about 2 months ago
Anthropic's Chloe Lubinski on AI, Ethics, and Future
Artificial Intelligence

Anthropic's Chloe Lubinski on AI, Ethics, and Future

Anthropic's Chloe Lubinski discusses AI's learning process, the importance of interpretability, and how AI reflects human values, emphasizing the need for ethical guidance in AI development.

about 2 months ago
AI Can't Prompt the Room, Says VisualLabs' Horváth
Artificial Intelligence

AI Can't Prompt the Room, Says VisualLabs' Horváth

Balázs Horváth of VisualLabs argues that AI cannot "prompt the room," emphasizing human judgment in defining what to build as the new bottleneck.

about 2 months ago
Isadora Martin-Dye: Layering AI Tone Instructions
Artificial Intelligence

Isadora Martin-Dye: Layering AI Tone Instructions

Isadora Martin-Dye explains why simple tone instructions for AI are insufficient, advocating for a four-layered approach to prompt engineering.

2 months ago
AI Agents in Healthcare Need Guardrails
Technology

AI Agents in Healthcare Need Guardrails

AI agents are transforming healthcare, but leaders must ask critical questions about data, governance, and control before widespread adoption.

2 months ago
OpenAI Simulates AI Deployments
Artificial Intelligence

OpenAI Simulates AI Deployments

OpenAI's new deployment simulation technique replays past conversations with candidate models to predict real-world behavior and mitigate risks before release.

2 months ago
Tejal Patwardhan: Stop Underestimating AI Models
Artificial Intelligence

Tejal Patwardhan: Stop Underestimating AI Models

Tejal Patwardhan of OpenAI discusses the evolution of AI evaluation, the concept of 'capability overhang,' and the need for realistic, real-world benchmarks.

2 months ago
Americans Want AI Cures, Fear Job Loss: Survey
Artificial Intelligence

Americans Want AI Cures, Fear Job Loss: Survey

A new survey reveals Americans prioritize AI for curing diseases but fear job loss and demand government regulation, showing broad consensus on AI's future.

2 months ago
OpenAI's Grand AI Ambition
Artificial Intelligence

OpenAI's Grand AI Ambition

OpenAI details its plan to democratize AI, ensuring AGI benefits all of humanity through broad access and shared prosperity.

3 months ago
OpenAI Launches AI Economic Research Program
Artificial Intelligence

OpenAI Launches AI Economic Research Program

OpenAI launches the Economic Research Exchange to fund external studies on AI's economic impacts, inviting researchers to collaborate using OpenAI tools.

3 months ago
Bengio: We're Building AI We Can't Control
AI Research

Bengio: We're Building AI We Can't Control

AI pioneer Yoshua Bengio warns that we are building increasingly powerful AI systems without fully understanding or controlling them, raising concerns about potential risks and the need for global safety standards.

3 months ago
OpenAI's AI Push Into Biodefense
Artificial Intelligence

OpenAI's AI Push Into Biodefense

OpenAI is now applying its advanced AI models, like GPT-Rosalind, to biodefense and pandemic preparedness, aiming to enhance global biological resilience.

3 months ago
AI: An Extension, Not a Replacement
AI Research

AI: An Extension, Not a Replacement

New research suggests AI's power lies in extending human cognition, not replicating it, impacting how we approach AI capabilities and safety.

3 months ago
AI Safety Pioneers: Tegmark & Esvelt on Guardrails
AI Research

AI Safety Pioneers: Tegmark & Esvelt on Guardrails

Max Tegmark and Kevin Esvelt discuss the critical importance of AI safety, the risks of advanced AI, and the need for global cooperation in shaping a beneficial future.

3 months ago
Anthropic's Olah on AI: Vatican Calls for Caution
Artificial Intelligence

Anthropic's Olah on AI: Vatican Calls for Caution

Anthropic co-founder Chris Olah addressed the Vatican's new AI encyclical, emphasizing the need for external critics and deeper societal discernment.

3 months ago
Pope Francis and AI: A Deep Dive into the Future
Artificial Intelligence

Pope Francis and AI: A Deep Dive into the Future

Pope Francis's first encyclical addresses Artificial Intelligence, signaling the Vatican's deep engagement with AI's opportunities and risks.

3 months ago
OpenAI bolsters AI content tracking
Artificial Intelligence

OpenAI bolsters AI content tracking

OpenAI is enhancing AI content tracking with C2PA conformance, Google SynthID watermarking, and a new public verification tool to boost transparency.

3 months ago
AI's 'Industrial Revolution' in Healthcare
Healthcare

AI's 'Industrial Revolution' in Healthcare

Boston Children's Hospital's Dr. Joan LaRovere discusses AI's transformative role in healthcare, emphasizing data diversity and personalized medicine.

3 months ago
AI Agents Flunk Social Reasoning Test
AI Research

AI Agents Flunk Social Reasoning Test

Microsoft's SocialReasoning-Bench reveals AI agents struggle to negotiate effectively in users' best interests, prioritizing task completion over optimal outcomes.

4 months ago
Personal AI: The New Personal Computer
Artificial Intelligence

Personal AI: The New Personal Computer

Explore the evolution from personal computers to personal AI, a shift promising dynamic, intelligent agents that understand context and proactively assist users.

4 months ago
OpenAI Taps New AI Talent
Artificial Intelligence

OpenAI Taps New AI Talent

OpenAI's 'ChatGPT Futures Class of 2026' honors 26 students using AI for ambitious projects, providing grants and access to advanced models.

4 months ago
Mozilla.ai: AI Sovereignty Beyond Borders
Technology

Mozilla.ai: AI Sovereignty Beyond Borders

Mozilla.ai's CEO John Dickerson redefines sovereign AI beyond geopolitics, emphasizing control, choice, and resilience at every level from nations to individuals.

4 months ago
Conditional Misalignment: A New AI Risk
AI Research

Conditional Misalignment: A New AI Risk

New research reveals that common LLM safety interventions fail under realistic data mixing, leading to conditional misalignment that standard evaluations miss.

4 months ago
OpenAI Faces Lawsuit Over Tumbler Ridge Shooting
Artificial Intelligence

OpenAI Faces Lawsuit Over Tumbler Ridge Shooting

Families sue OpenAI after the Tumbler Ridge shooting, alleging the company ignored ChatGPT warnings from the attacker.

4 months ago
AI Agents Lack Identity, Risking Enterprise Trust
Technology

AI Agents Lack Identity, Risking Enterprise Trust

Enterprises are struggling with the AI agent identity problem, a critical gap in governance and accountability that hinders trust and adoption.

4 months ago
OpenAI's Guiding Principles for AGI
Artificial Intelligence

OpenAI's Guiding Principles for AGI

OpenAI outlines its guiding principles for AGI development, emphasizing democratization, empowerment, universal prosperity, resilience, and adaptability.

4 months ago
Claude's 2026 Election Safeguards
Artificial Intelligence

Claude's 2026 Election Safeguards

Anthropic details its 2026 election safeguards for Claude, focusing on bias mitigation, policy enforcement, and providing users with reliable, up-to-date information.

4 months ago
AI's Data Problem: More Isn't Always Better
Artificial Intelligence

AI's Data Problem: More Isn't Always Better

Janusz Marecki, CEO of Fractal Brain, discusses the limitations of current AI models and the shift towards data quality and specialized techniques like synthetic data.

4 months ago
OpenAI's Huet on Building AI for Connection, Not Just Code
Artificial Intelligence

OpenAI's Huet on Building AI for Connection, Not Just Code

OpenAI's Romain Huet and Hearth AI founder Ashe Magalhaes discuss building user-centric AI, fostering connection, and the future of AI development.

5 months ago
Simon Podhajsky on "Cognitive Exhaust Fumes"
Artificial Intelligence

Simon Podhajsky on "Cognitive Exhaust Fumes"

Simon Podhajsky discusses 'Cognitive Exhaust Fumes,' advocating for read-only AI observers to analyze personal data and reveal cognitive patterns, contrasting this with riskier AI agents.

5 months ago
IBM Experts on AI Ethics & Autonomous Systems
Artificial Intelligence

IBM Experts on AI Ethics & Autonomous Systems

IBM AI experts Sandi Besen and Gabe Goodhart discuss AI ethics in autonomous systems, cognitive offloading, and the future of human-AI collaboration.

5 months ago
OpenAI's Blueprint for AI Behavior
Artificial Intelligence

OpenAI's Blueprint for AI Behavior

OpenAI unveils its formal Model Spec, a public framework detailing intended AI behavior and a 'Chain of Command' for resolving conflicting instructions.

5 months ago
IBM's Martin Keen on AI Human-in-the-Loop Spectrum
Artificial Intelligence

IBM's Martin Keen on AI Human-in-the-Loop Spectrum

IBM's Martin Keen explains the human-in-the-loop spectrum for AI, detailing how human involvement is crucial in training, tuning, and inference stages.

5 months ago
AI Governance: Control, Not Code, Drives Success
Technology

AI Governance: Control, Not Code, Drives Success

Enterprise AI success hinges on robust governance, focusing on control and trust rather than just code, as Databricks leaders explain.

6 months ago
Anthropic Launches AI Futures Think Tank
Artificial Intelligence

Anthropic Launches AI Futures Think Tank

Anthropic launches The Anthropic Institute to research and address the societal challenges posed by advanced AI development.

6 months ago
Reasoning Nudges LLMs Towards Honesty
AI Research

Reasoning Nudges LLMs Towards Honesty

New research reveals that LLM reasoning enhances honesty not through content, but by leveraging the geometry of representational spaces, stabilizing honest defaults.

6 months ago
AI Agents Need Humans: The HITL Advantage
Artificial Intelligence

AI Agents Need Humans: The HITL Advantage

IBM AI Engineer Anna Gutowska explains why human intervention in AI agents is critical for preventing subtle errors and ensuring safe, effective deployment.

6 months ago
Qasar Younis on AI's Future Impact
Artificial Intelligence

Qasar Younis on AI's Future Impact

Applied Intuition CEO Qasar Younis discusses the transformative impact of AI on industries, the importance of understanding the technology, and the future of autonomous systems.

6 months ago