1 articles with this tag
DigitalOcean's Archana Kamath and Tyler Gillam discuss model routing, arguing that preferences like cost and latency should dictate LLM choices over benchmarks, and showcase their open-source inference router.