AI News HubLIVE
站内改写1 分钟阅读

待翻译:Nvidia's Switchyard router reshuffles AI models mid-task to cut task costs

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks to cheaper models and that becomes its…

来源Hacker News AI作者: gmays

AI 服务暂时不可用,以下为来源正文,待恢复后补全翻译。

Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks to cheaper models and that becomes its own engineering project, one that has to be maintained every time a workflow changes. Nvidia is proposing a fix that touches both ends of that problem at once. The company is out on Tuesday with Nemotron 3.5 Lightning, a 30-billion-parameter open mixture-of-experts model built for high-volume, specialized agent tasks, alongside NeMo Switchyard, an open-source library that routes each step of an agent workflow to whichever model fits it best.