#AI Coding

13 articles with this tag

SpaceX paid $60B for the editor. Four other rounds this week completed the factory it sits inside.
AI News

SpaceX paid $60B for the editor. Four other rounds this week completed the factory it sits inside.

River AI, Lovable, CodeRabbit, and Blacksmith raised $1.7B to form a closed AI coding loop. Then SpaceX paid $60B for Cursor to own the interface layer.

2 days ago
Claude's Corner: Canary - The AI QA Engineer That Reads Your Code
Claude's Corner

Claude's Corner: Canary - The AI QA Engineer That Reads Your Code

Canary reads your source code, understands your application's intent, and automatically generates and runs end-to-end browser tests when you open a pull request. Founded by ex-Windsurf and Google engineers, it is the clearest bet yet that AI coding tools need an AI QA counterpart.

16 days ago
Kimi K2.6 Open Sources Advanced Coding AI
AI

Kimi K2.6 Open Sources Advanced Coding AI

Moonshot AI open-sources Kimi K2.6, a powerful AI model for coding and agentic workflows, boasting state-of-the-art long-horizon execution and agent swarm capabilities.

about 1 month ago
AI Code Generators Spark 'Bug Apocalypse,' Experts Warn
Cybersecurity

AI Code Generators Spark 'Bug Apocalypse,' Experts Warn

Jack Cable of Corridor discusses the 'AI bug apocalypse,' the increasing vulnerability of software due to AI coding tools, and the need for secure-by-design principles.

about 1 month ago
OpenAI Flags Major Flaws in SWE-Bench Pro
Artificial Intelligence

OpenAI Flags Major Flaws in SWE-Bench Pro

OpenAI's audit reveals approximately 30% of SWE-Bench Pro's coding tasks are flawed, prompting the company to retract its recommendation for the benchmark.

about 1 month ago
WorkOS's Zack Proser on Untethered AI Productivity
Artificial Intelligence

WorkOS's Zack Proser on Untethered AI Productivity

WorkOS's Zack Proser discusses how developers can maintain productivity and well-being amidst the rapid advancement of AI coding agents, focusing on balance and intentional workflows.

2 months ago
DeepSeek V4 vs. Opus: Ahmad Awais on AI Coding Taste
Artificial Intelligence

DeepSeek V4 vs. Opus: Ahmad Awais on AI Coding Taste

Ahmad Awais discusses how AI coding agents can learn 'coding taste' to outperform generic models, focusing on the difference between functional code and good design.

2 months ago
Sarah Chieng: Fast Models Need "Slow" Developers
Artificial Intelligence

Sarah Chieng: Fast Models Need "Slow" Developers

Cerebras' Sarah Chieng discusses how fast AI coding models like Codex Spark necessitate new developer habits and workflows for optimal results.

3 months ago
ClickUp's 22% cut comes with $1M salary bands. Evans calls it the 100x org.
Artificial Intelligence

ClickUp's 22% cut comes with $1M salary bands. Evans calls it the 100x org.

ClickUp CEO Zeb Evans announced a 22% headcount cut and the introduction of $1M cash salary bands, framed not as cost-cutting but as a restructure around what he calls the 100x organization. His diagnosis of how AI is reshaping engineering org design is sharper than most peer SaaS messaging.

3 months ago
Cielara Code Outperforms Rivals
AI Research

Cielara Code Outperforms Rivals

Cielara Code, from Causal Dynamics Lab, significantly improves AI coding agent performance by mapping production software, outperforming rivals in key benchmarks.

3 months ago
OpenAI Unveils GPT-5.5
Artificial Intelligence

OpenAI Unveils GPT-5.5

OpenAI launches GPT-5.5, boasting enhanced intelligence, autonomy, and speed for complex tasks, alongside advanced safety features.

4 months ago
AI Coding Benchmark Scores Skewed by Infrastructure
Artificial Intelligence

AI Coding Benchmark Scores Skewed by Infrastructure

Infrastructure configuration, not just AI model prowess, can significantly skew benchmark results, complicating deployment decisions.

5 months ago
Claude Opus 4.6: Smarter, Faster, and Longer Context
Artificial Intelligence

Claude Opus 4.6: Smarter, Faster, and Longer Context

Anthropic's Claude Opus 4.6 launches with a 1M token context window, enhanced coding, and state-of-the-art benchmark performance.

6 months ago