跳到主要内容
AI News HubLIVE
公开文章 20采集文章 21可信度 84刷新频率 120 分钟
健康状态 健康来源类型 官方原文权限 官方原文最近入库 2026-08-24ID groq-blog运行状态 已启用

Official AI inference platform blog; confirm reuse terms before full body display.

最新公开文章

待翻译:Groq Among the First to Bring NVIDIA Groq 3 LPX and Vera Rubin NVL72 to Market

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:We are thrilled to announce that Groq will be among the first adopters of NVIDIA Groq 3 LPX, boosting inference token generation for NVIDIA Vera Rubin NVL72 which it will deploy to its purpose-built AI inference cloud.…

Groq Blog站内正文待翻译:Groq Among the First to Bring NVIDIA Groq 3 LPX and Vera Rubin NVL72 to Market

待翻译:Groq Launches Meta's Llama 3 Instruct AI Models on LPU™ Inference Engine

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Llama 3 Now Available to Developers via GroqChat and GroqCloud™ Here’s what’s happened in the last 36 hours: April 18th, Noon: Meta releases versions of its latest Large Language Model (LLM), Llama 3. April 19th, Midnig…

Groq Blog站内正文待翻译:Groq Launches Meta's Llama 3 Instruct AI Models on LPU™ Inference Engine

待翻译:Groq Partners with Aramco on World’s Largest AI Data Center

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Jonathan Ross, CEO and Founder of Groq, and Tareq Amin, CEO of Aramco Digital, a subsidiary of Saudi Aramco, proudly announced their partnership to build the largest AI inference data center in Saudi Arabia. The data ce…

Groq Blog站内正文待翻译:Groq Partners with Aramco on World’s Largest AI Data Center

待翻译:Context Length in LLMs: Optimize Business AI Performance

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:What is Context Length? As businesses look to leverage Large Language Models (LLMs) for conversational AI, generative AI, and analytics, a crucial factor often gets overlooked: context length, also known as context wind…

Groq Blog站内正文待翻译:Context Length in LLMs: Optimize Business AI Performance

待翻译:The Five Future Stages of Generative AI

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:This blog is adapted from an original post by Groq CEO and Founder, Jonathan Ross. If Large Language Models (LLMs) are the printing press of the Generative AI age, then what’s next? What will be the AI equivalent of the…

Groq Blog站内正文待翻译:The Five Future Stages of Generative AI

待翻译:What is AI Inference? ML Basics Explained

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The rapidly evolving field of Artificial Intelligence (AI) has led to significant advancements in Machine Learning (ML), with "inference" emerging as a crucial concept. But what exactly is inference, and how does it wor…

Groq Blog站内正文待翻译:What is AI Inference? ML Basics Explained

待翻译:Thank You! 1 Million Developers Now On GroqCloud™

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Today, one year this week since we launched GroqCloud, we're celebrating the one million developers who are now on GroqCloud! In just a year this community of builders, makers, and innovators have shipped incredible app…

Groq Blog站内正文待翻译:Thank You! 1 Million Developers Now On GroqCloud™

待翻译:What is a Language Processing Unit?

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Overview Groq LPU™ AI Inference Technology Groq builds fast AI inference. Groq® LPU™ AI inference technology delivers exceptional AI compute speed, quality, and affordability at scale. Groq AI inference infrastructure,…

Groq Blog站内正文待翻译:What is a Language Processing Unit?

待翻译:Batch Processing with GroqCloud™ for AI Inference Workloads

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:GroqCloud™ provides fast inference for complex AI solutions that require instant responsiveness. But what happens when your use cases expand and require features beyond speed? That’s where Batch Processing comes in –now…

Groq Blog站内正文待翻译:Batch Processing with GroqCloud™ for AI Inference Workloads

待翻译:From Speed to Scale: How Groq Is Optimized for MoE & Other Large Models

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:You know Groq runs small models. But did you know we run large models including MoE uniquely well? Here’s why. The Evolution of Advanced Openly-Available LLMs There’s no argument that Artificial intelligence (AI) has ex…

Groq Blog站内正文待翻译:From Speed to Scale: How Groq Is Optimized for MoE & Other Large Models

待翻译:Groq Names Simon Edwards Chief Financial Officer

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Appointment supports Groq’s global expansion and growing AI inference demand Mountain View, Calif. — September 22, 2025 — Groq, the global pioneer in AI inference, today announced the appointment of Simon Edwards as Chi…

Groq Blog站内正文待翻译:Groq Names Simon Edwards Chief Financial Officer

GroqCloud Beta版推出远程MCP支持

GroqCloud宣布其远程模型上下文协议(MCP)服务器集成功能已进入Beta阶段,开发者可无缝连接外部工具,实现更快、更低成本的AI应用。该功能兼容OpenAI API,支持零代码迁移。

Groq Blog站内正文GroqCloud Beta版推出远程MCP支持

GroqCloud为GPT-OSS模型推出提示缓存与降价措施

Groq宣布对其GPT-OSS模型进行两项重要更新:降低价格和推出提示缓存功能,旨在提升AI推理的成本效益和速度。降价立即生效,并追溯至2025年10月所有未付款发票。提示缓存可带来高达50%的缓存令牌折扣、更低的延迟以及更高的速率限制,且无需任何配置。

Groq Blog站内正文GroqCloud为GPT-OSS模型推出提示缓存与降价措施

产品内集成LLM:实用现场指南

本文基于实践经验,介绍如何将开源LLM可靠地集成到产品中。核心是四步循环:读取(仅取必要上下文)、约束(明确系统和格式规则)、执行(结构化输出、函数调用或纯文本)、解释(向用户展示步骤和引用)。还涵盖常见模式(路由器、提取器、翻译器等)、安全发布(测试、监控、回退)及常见陷阱。目标是打造用户无感知、可靠的AI特性。

Groq Blog站内正文产品内集成LLM:实用现场指南

OpenAI 开放安全模型首发支持

GroqCloud 宣布即日起支持 OpenAI 最新开源安全模型 GPT-OSS-Safeguard-20B,提供超过 1000 t/s 的推理速度。该模型专为安全分类工作负载设计,支持用户自定义策略、可配置推理力度及完整推理轨迹,适用于企业文档扫描、AI 聊天机器人、政策审计和用户生成内容平台等场景。定价与基础 GPT-OSS-20B 相同,输入 token $0.075/M,输出 token $0.30/M。

Groq Blog站内正文OpenAI 开放安全模型首发支持

GroqCloud 推出远程 MCP 支持测试版

Groq 宣布在 GroqCloud 上推出 MCP 连接器测试版,率先支持 Google Workspace(Gmail、云端硬盘和日历)。这些预建的 MCP 服务器由 Groq 托管,使 AI 代理能够通过 Responses API 与 Google 工具交互,而无需管理自己的 MCP 服务器。

Groq Blog站内正文GroqCloud 推出远程 MCP 支持测试版

Groq 被 2025 年 Gartner® AI 基础架构酷供应商报告收录

Groq 凭借其 LPU 芯片的确定性、低延迟推理和线性扩展能力,被 Gartner 评为 2025 年 AI 基础架构领域的酷供应商。超过 250 万开发者使用 Groq,其性能比 GPU 快 5 倍且成本更低。

Groq Blog站内正文Groq 被 2025 年 Gartner® AI 基础架构酷供应商报告收录

推动美国人工智能堆栈发展

文章讨论了美国在人工智能计算领域的领导地位,特别是推理计算的重要性,以及如何通过出口政策维持优势。强调了市场驱动的生态系统和行业联盟的作用,建议采用灵活的多模型框架。

Groq Blog站内正文推动美国人工智能堆栈发展

GroqCloud:扩展以满足需求

GroqCloud正在全球扩展其AI推理基础设施,以应对实时应用从实验转向生产带来的需求增长。最近在英国新建的数据中心,与Equinix合作,为欧洲开发者和企业提供低延迟、高性能的推理服务。GroqCloud现已拥有超过350万开发者,生产流量持续增长。

Groq Blog站内正文GroqCloud:扩展以满足需求

深度解析 LPU:Groq 速度背后的秘密

Groq 的 LPU 是专为推理设计的硬件,通过 TruePoint 数字、SRAM 存储、静态调度和实时张量并行等技术,在不牺牲精度的情况下实现超低延迟推理。Moonshot 的 Kimi K2 模型在 Groq 上以 40 倍性能运行,展示了 LPU 架构的优势。

Groq Blog站内正文深度解析 LPU:Groq 速度背后的秘密

全部来源