AI News HubLIVE
站內改寫1 分鐘閱讀

待翻譯:Nvidia's Switchyard router reshuffles AI models mid-task to cut task costs

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks to cheaper models and that becomes its…

來源Hacker News AI作者: gmays

AI 服務暫時不可用,以下為來源正文,待恢復後補全翻譯。

Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks to cheaper models and that becomes its own engineering project, one that has to be maintained every time a workflow changes. Nvidia is proposing a fix that touches both ends of that problem at once. The company is out on Tuesday with Nemotron 3.5 Lightning, a 30-billion-parameter open mixture-of-experts model built for high-volume, specialized agent tasks, alongside NeMo Switchyard, an open-source library that routes each step of an agent workflow to whichever model fits it best.