#Coding Agents

20 articles with this tag

OpenCode CEO on 20x Growth & AI Agent Market
Startup News

OpenCode CEO on 20x Growth & AI Agent Market

OpenCode CEO Jay V reveals how the platform achieved 20x growth, reaching 4.6M users by supporting any AI model and becoming a key player in the global coding agent market.

24 days ago
RLM Models: A New Approach to Large Codebases
Artificial Intelligence

RLM Models: A New Approach to Large Codebases

Shashi Jagtap of Superagentic AI explores Recursive Language Models (RLMs) and their application for coding agents tackling large codebases, showcasing their RLM Code implementation.

about 1 month ago
Databricks Benchmarks AI Coding Tools
Technology

Databricks Benchmarks AI Coding Tools

Databricks benchmarks AI coding agents on its multi-million line codebase, finding open-source models competitive and token price an unreliable cost indicator.

about 1 month ago
Teaching AI Agents to Master Spreadsheets
Artificial Intelligence

Teaching AI Agents to Master Spreadsheets

Nuno Campos of Witan Labs discusses teaching AI agents to master spreadsheets using a REPL approach, improving accuracy and efficiency.

about 1 month ago
AI Agents Don't Always Follow Rules
Artificial Intelligence

AI Agents Don't Always Follow Rules

Talha Sheikh from Checkout.com discusses the unreliability of AI coding agents and the critical need for verification layers and guardrails to ensure dependable AI outputs.

about 1 month ago
Andrew Dumit on "Respect The Process" at Watershed
Artificial Intelligence

Andrew Dumit on "Respect The Process" at Watershed

Andrew Dumit from Watershed discusses how to build trustworthy AI coding agents by respecting the process and implementing deterministic execution.

about 1 month ago
SWE-Marathon: Evaluating AI Coding Agents at Scale
AI Research

SWE-Marathon: Evaluating AI Coding Agents at Scale

Rishi Desai from Abundant AI introduces SWE-Marathon, a benchmark evaluating AI coding agents on billion-token scale tasks, revealing current limitations and the need for robust verification.

about 1 month ago
Evaluating Coding Agents: Lessons from SWE-rebench
AI Research

Evaluating Coding Agents: Lessons from SWE-rebench

Ibragim Badertdinov from Nebius shares key lessons from evaluating coding agents using the SWE-rebench benchmark, highlighting the importance of real-world tasks, reliable verification, and cost-effectiveness.

2 months ago
Devin's 80% Moment: AI Coding Agents Evolve
Artificial Intelligence

Devin's 80% Moment: AI Coding Agents Evolve

Walden Yan and Cole Murray discuss Devin's '80% moment' in AI coding, highlighting background agents, multiple PRs, and the end of hand-held coding.

3 months ago
Hugging Face's Ben Burtenshaw on AI System Engineering
Artificial Intelligence

Hugging Face's Ben Burtenshaw on AI System Engineering

Ben Burtenshaw from Hugging Face discusses how AI coding agents can be used for AI system engineering, kernel optimization, and building multi-agent autoresearch labs.

3 months ago
Coding Agent Inference Benchmark Revealed
Technology

Coding Agent Inference Benchmark Revealed

Together AI unveils a new benchmark for coding agent inference, highlighting performance under real-world load and significant cost advantages.

3 months ago
Marlene Mhangami: Playwright for Functionality Testing
Technology

Marlene Mhangami: Playwright for Functionality Testing

Marlene Mhangami from Microsoft and GitHub discusses leveraging Playwright and AI agents for effective functionality testing, emphasizing clean code and behavior-driven development.

3 months ago
OpenAI's "Parameter Golf" Reveals AI's Role
Artificial Intelligence

OpenAI's "Parameter Golf" Reveals AI's Role

OpenAI's "Parameter Golf" competition revealed how AI coding agents are transforming machine learning research, pushing innovation under tight constraints.

3 months ago
VIBE✓ adds friction to AI coding agents
Technology

VIBE✓ adds friction to AI coding agents

Mozilla.ai's VIBE✓ framework introduces deliberate friction to coding agent workflows, mitigating automation bias and ensuring human oversight.

3 months ago
Embedding OpenClaw Coding Agent in Your Product
Artificial Intelligence

Embedding OpenClaw Coding Agent in Your Product

Matthias Luebken from Tavon.ai discusses embedding the OpenClaw coding agent, Pi, into products, highlighting its utility for developers and the future of AI in software systems.

3 months ago
OpenAI's Safety Playbook for Codex
Artificial Intelligence

OpenAI's Safety Playbook for Codex

OpenAI details its robust safety measures for its Codex AI coding agent, emphasizing sandboxing, network controls, and detailed telemetry for secure deployment.

3 months ago
Databricks Tames Coding AI Chaos
Technology

Databricks Tames Coding AI Chaos

Databricks introduces Unity AI Gateway to manage AI coding agents, offering centralized governance, cost controls, and observability for enterprises.

4 months ago
Databricks Centralizes Coding AI
Technology

Databricks Centralizes Coding AI

Databricks launches AI Gateway to centralize governance, security, and cost controls for the growing number of AI coding agents used by enterprises.

4 months ago
Exa Unveils New Code Search Benchmarks
Artificial Intelligence

Exa Unveils New Code Search Benchmarks

Exa.ai releases 'WebCode', a new benchmark suite for evaluating search performance in coding agents, addressing limitations in existing tools.

5 months ago
AI Agents Leveled Up by Harness Engineering
Artificial Intelligence

AI Agents Leveled Up by Harness Engineering

LangChain's harness engineering approach dramatically improved an AI coding agent's performance by refining its surrounding system, not the core model.

6 months ago