Why Decagon Runs 90 Percent of Its Enterprise AI on Open Source

Decagon founders Jesse Zhang and Ashwin Sreenivas explain why fine-tuned open-source AI beats frontier APIs on speed, cost, and enterprise performance.

8 min read
Decagon co-founders Jesse Zhang and Ashwin Sreenivas discussing enterprise AI playbooks
Decagon co-founders discuss open-source model optimization and enterprise AI strategy.· a16z

Visual TL;DR. Initial Frontier API Use led to Scaling Latency Issues. Scaling Latency Issues drove Open Source Adoption. Open Source Adoption enables Fine-Tuning Control. Open Source Adoption yields Performance Gains. Performance Gains supports Software Layer Moats. Performance Gains informs Redefine Engineering. Performance Gains demonstrates Enterprise AI Success.

  1. Initial Frontier API Use: Decagon initially relied on closed-source models for quick product launch
  2. Scaling Latency Issues: handling millions of interactions and voice agents made latency an existential metric
  3. Open Source Adoption: 90% of Decagon's platform now runs on fine-tuned open-source models
  4. Fine-Tuning Control: open-source models offer control over smaller models, unlike frontier labs
  5. Performance Gains: fine-tuned open-source AI beats frontier APIs on speed, cost, and enterprise performance
  6. Software Layer Moats: true moats are built on software layers, not just calling frontier APIs
  7. Redefine Engineering: shifting from API calls to deep model integration redefines forward deployment
  8. Enterprise AI Success: Decagon's operational reality shows open source is key for enterprise AI
Visual TL;DR
Visual TL;DR, startuphub.ai Scaling Latency Issues drove Open Source Adoption. Open Source Adoption yields Performance Gains. Performance Gains demonstrates Enterprise AI Success drove yields demonstrates Scaling Latency Issues Open Source Adoption Performance Gains Enterprise AI Success From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Scaling Latency Issues drove Open Source Adoption. Open Source Adoption yields Performance Gains. Performance Gains demonstrates Enterprise AI Success drove yields demonstrates Scaling LatencyIssues Open SourceAdoption Performance Gains Enterprise AISuccess From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Scaling Latency Issues drove Open Source Adoption. Open Source Adoption yields Performance Gains. Performance Gains demonstrates Enterprise AI Success drove yields demonstrates Scaling Latency Issues handling millions of interactions andvoice agents made latency an existentialmetric Open Source Adoption 90% of Decagon's platform now runs onfine-tuned open-source models Performance Gains fine-tuned open-source AI beats frontierAPIs on speed, cost, and enterpriseperformance Enterprise AI Success Decagon's operational reality shows opensource is key for enterprise AI From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Scaling Latency Issues drove Open Source Adoption. Open Source Adoption yields Performance Gains. Performance Gains demonstrates Enterprise AI Success drove yields demonstrates Scaling LatencyIssues handling millionsof interactions andvoice agents made… Open SourceAdoption 90% of Decagon'splatform now runson fine-tuned… Performance Gains fine-tunedopen-source AIbeats frontier APIs… Enterprise AISuccess Decagon'soperational realityshows open source… From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Initial Frontier API Use led to Scaling Latency Issues. Scaling Latency Issues drove Open Source Adoption. Open Source Adoption enables Fine-Tuning Control. Open Source Adoption yields Performance Gains. Performance Gains supports Software Layer Moats. Performance Gains informs Redefine Engineering. Performance Gains demonstrates Enterprise AI Success led to drove enables yields supports informs demonstrates Initial Frontier API Use Decagon initially relied on closed-sourcemodels for quick product launch Scaling Latency Issues handling millions of interactions andvoice agents made latency an existentialmetric Open Source Adoption 90% of Decagon's platform now runs onfine-tuned open-source models Fine-Tuning Control open-source models offer control oversmaller models, unlike frontier labs Performance Gains fine-tuned open-source AI beats frontierAPIs on speed, cost, and enterpriseperformance Software Layer Moats true moats are built on software layers,not just calling frontier APIs Redefine Engineering shifting from API calls to deep modelintegration redefines forward deployment Enterprise AI Success Decagon's operational reality shows opensource is key for enterprise AI From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Initial Frontier API Use led to Scaling Latency Issues. Scaling Latency Issues drove Open Source Adoption. Open Source Adoption enables Fine-Tuning Control. Open Source Adoption yields Performance Gains. Performance Gains supports Software Layer Moats. Performance Gains informs Redefine Engineering. Performance Gains demonstrates Enterprise AI Success led to drove enables yields supports informs demonstrates Initial FrontierAPI Use Decagon initiallyrelied onclosed-source… Scaling LatencyIssues handling millionsof interactions andvoice agents made… Open SourceAdoption 90% of Decagon'splatform now runson fine-tuned… Fine-TuningControl open-source modelsoffer control oversmaller models,… Performance Gains fine-tunedopen-source AIbeats frontier APIs… Software LayerMoats true moats arebuilt on softwarelayers, not just… RedefineEngineering shifting from APIcalls to deep modelintegration… Enterprise AISuccess Decagon'soperational realityshows open source… From startuphub.ai · The publishers behind this format

In the rapid rush to deploy enterprise AI agents, many founders assumed the primary battle would be won by calling frontier APIs like OpenAI and Anthropic. But on a recent episode of The a16z Show, Decagon co-founders Jesse Zhang and Ashwin Sreenivas outlined a starkly different operational reality: 90 percent of their platform workflow now runs on fine-tuned open-source models.

Why Decagon Runs 90 Percent of Its Enterprise AI on Open Source - a16z
Why Decagon Runs 90 Percent of Its Enterprise AI on Open Source — from a16z

The Shift from Frontier APIs to Fine-Tuned Open Source

When Decagon launched its enterprise AI customer service platform, the engineering team relied almost exclusively on frontier closed-source models. The priority was getting a functional product to market quickly. However, as the company scaled to handle millions of customer interactions for global brands and launched voice agents, latency became an existential metric.

"When you want to go to smaller models, unfortunately, the frontier labs do have small models, but you can't really control them in the way that you want," Zhang explained during the interview. "Most small models out of the box are not going to be good enough at the task that we want them to do. So you have to fine-tune them."

By breaking complex agentic workflows into distinct, discrete sub-tasks, such as topic classification or fraud detection, Decagon realized that individual steps do not require general intelligence like coding or advanced mathematics. A smaller, open-source model trained specifically on that single task delivers equal or superior accuracy with significantly lower latency.

The False Trade-Off Between Intelligence, Cost, and Speed

A common belief in Silicon Valley holds that teams must choose between high-cost frontier intelligence and cheaper, dumber models. Sreenivas argued that this framing fundamentally misinterprets post-training capabilities in specialized enterprise domains.

"When we fine-tune smaller, dumber models, it's that they're just not as general purpose, but on the specific task we want them to do, they actually outperform the large, smart, state-of-the-art models," Sreenivas said. "So we end up getting all three things. It is better at the task, it is cheaper, and it is faster."

Decagon still reserves closed frontier models for the remaining 10 percent of its workload. These include open-ended, exploratory tasks such as Duet Autopilot, an auxiliary agent that analyzes millions of historical customer transcripts, identifies recurring failure patterns, and automatically drafts process updates and test simulations.

Why Software Layer Moats Persist in the Age of AGI

The conversation addressed the industry debate over whether general artificial intelligence will commoditize application software, turning application startups into thin user interfaces paired with manual implementation staff. Both founders pushed back against the narrative that frontier labs will absorb the entire software market.

Even if foundational model intelligence reaches general human capability, enterprise deployments require extensive software infrastructure around the model. Large organizations need granular permissioning, audit trails, system integrations, compliance monitoring, and mechanisms to encode complex business logic.

StartupHub.ai data gives Decagon a platform score of 65/100 as it competes in the enterprise intelligence sector. Among broader general intelligence startups tracked by StartupHub.ai data, AGI holds a score of 47/100 with $10M raised in verified seed funding, alongside peers such as Adept AI (71/100), Hebbia (72/100), Peak (70/100), Inflection AI (55/100), and Manus AI (75/100).

Redefining the Forward Deployed Engineering Model

With forward deployed engineers becoming a popular hiring trend across Silicon Valley, Sreenivas, a former deployment strategist at Palantir Technologies (NYSE:PLTR), warned that mismanaging the role turns software startups into service firms.

Because AI workflows are novel, forward deployed teams are initially essential to embed with customers and uncover how work actually gets done. However, those learnings must immediately flow back into the core product architecture.

"Forward deployed engineers eat pain and excrete product," Sreenivas noted, quoting an internal Palantir maxim. He emphasized that if forward deployed engineers spend their time writing custom, one-off code for individual client requests rather than building scalable platform features, the business risks becoming a glorified IT consultancy akin to Accenture (NYSE:ACN) rather than a scalable software business.

© 2026 StartupHub.ai. All rights reserved. Do not enter, scrape, copy, reproduce, or republish this article in whole or in part. Use as input to AI training, fine-tuning, retrieval-augmented generation, or any machine-learning system is prohibited without written license. Substantially-similar derivative works will be pursued to the fullest extent of applicable copyright, database, and computer-misuse laws. See our terms.