本文にスキップ
AI News HubLIVE
公開記事 20収集記事 21信頼度 84更新頻度 120 分
稼働状態 正常ソース種別 公式全文利用権限 公式全文最終取り込み 2026-08-24ID groq-blog状態 有効

Official AI inference platform blog; confirm reuse terms before full body display.

最新公開記事

翻訳待ち:Groq Among the First to Bring NVIDIA Groq 3 LPX and Vera Rubin NVL72 to Market

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:We are thrilled to announce that Groq will be among the first adopters of NVIDIA Groq 3 LPX, boosting inference token generation for NVIDIA Vera Rubin NVL72 which it will deploy to its purpose-built AI inference cloud.…

Groq Blogサイト内本文翻訳待ち:Groq Among the First to Bring NVIDIA Groq 3 LPX and Vera Rubin NVL72 to Market

翻訳待ち:Groq Launches Meta's Llama 3 Instruct AI Models on LPU™ Inference Engine

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Llama 3 Now Available to Developers via GroqChat and GroqCloud™ Here’s what’s happened in the last 36 hours: April 18th, Noon: Meta releases versions of its latest Large Language Model (LLM), Llama 3. April 19th, Midnig…

Groq Blogサイト内本文翻訳待ち:Groq Launches Meta's Llama 3 Instruct AI Models on LPU™ Inference Engine

翻訳待ち:Groq Partners with Aramco on World’s Largest AI Data Center

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Jonathan Ross, CEO and Founder of Groq, and Tareq Amin, CEO of Aramco Digital, a subsidiary of Saudi Aramco, proudly announced their partnership to build the largest AI inference data center in Saudi Arabia. The data ce…

Groq Blogサイト内本文翻訳待ち:Groq Partners with Aramco on World’s Largest AI Data Center

翻訳待ち:Context Length in LLMs: Optimize Business AI Performance

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:What is Context Length? As businesses look to leverage Large Language Models (LLMs) for conversational AI, generative AI, and analytics, a crucial factor often gets overlooked: context length, also known as context wind…

Groq Blogサイト内本文翻訳待ち:Context Length in LLMs: Optimize Business AI Performance

翻訳待ち:The Five Future Stages of Generative AI

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:This blog is adapted from an original post by Groq CEO and Founder, Jonathan Ross. If Large Language Models (LLMs) are the printing press of the Generative AI age, then what’s next? What will be the AI equivalent of the…

Groq Blogサイト内本文翻訳待ち:The Five Future Stages of Generative AI

翻訳待ち:What is AI Inference? ML Basics Explained

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:The rapidly evolving field of Artificial Intelligence (AI) has led to significant advancements in Machine Learning (ML), with "inference" emerging as a crucial concept. But what exactly is inference, and how does it wor…

Groq Blogサイト内本文翻訳待ち:What is AI Inference? ML Basics Explained

翻訳待ち:Thank You! 1 Million Developers Now On GroqCloud™

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Today, one year this week since we launched GroqCloud, we're celebrating the one million developers who are now on GroqCloud! In just a year this community of builders, makers, and innovators have shipped incredible app…

Groq Blogサイト内本文翻訳待ち:Thank You! 1 Million Developers Now On GroqCloud™

翻訳待ち:What is a Language Processing Unit?

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Overview Groq LPU™ AI Inference Technology Groq builds fast AI inference. Groq® LPU™ AI inference technology delivers exceptional AI compute speed, quality, and affordability at scale. Groq AI inference infrastructure,…

Groq Blogサイト内本文翻訳待ち:What is a Language Processing Unit?

翻訳待ち:Batch Processing with GroqCloud™ for AI Inference Workloads

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:GroqCloud™ provides fast inference for complex AI solutions that require instant responsiveness. But what happens when your use cases expand and require features beyond speed? That’s where Batch Processing comes in –now…

Groq Blogサイト内本文翻訳待ち:Batch Processing with GroqCloud™ for AI Inference Workloads

翻訳待ち:From Speed to Scale: How Groq Is Optimized for MoE & Other Large Models

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:You know Groq runs small models. But did you know we run large models including MoE uniquely well? Here’s why. The Evolution of Advanced Openly-Available LLMs There’s no argument that Artificial intelligence (AI) has ex…

Groq Blogサイト内本文翻訳待ち:From Speed to Scale: How Groq Is Optimized for MoE & Other Large Models

翻訳待ち:Groq Names Simon Edwards Chief Financial Officer

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Appointment supports Groq’s global expansion and growing AI inference demand Mountain View, Calif. — September 22, 2025 — Groq, the global pioneer in AI inference, today announced the appointment of Simon Edwards as Chi…

Groq Blogサイト内本文翻訳待ち:Groq Names Simon Edwards Chief Financial Officer

GroqCloudでリモートMCPサポートがベータ版に

GroqCloudは、リモートモデルコンテキストプロトコル(MCP)サーバー統合のベータ版提供を発表。OpenAI互換のAPIを介して外部ツールに接続でき、コード変更なしで高速かつ低コストなAIアプリケーションを実現します。

Groq Blogサイト内本文GroqCloudでリモートMCPサポートがベータ版に

GroqCloud、GPT‑OSSモデルにプロンプトキャッシングと値下げを導入

Groqは、GPT-OSSモデルに対する2つの重要なアップデート、価格引き下げとプロンプトキャッシング機能の提供を発表しました。これらはAI推論のコスト効率と速度を向上させることを目的としています。新価格は即時発効し、2025年10月の未払い請求書にも遡及適用されます。プロンプトキャッシングにより、キャッシュされたトークンが最大50%割引、レイテンシが低減、レート制限が緩和され、設定は不要です。

Groq Blogサイト内本文GroqCloud、GPT‑OSSモデルにプロンプトキャッシングと値下げを導入

プロダクト内でのLLM活用:実践的フィールドガイド

本稿は、オープンソースLLMを実際のプロダクトに確実に統合するための実践的ガイドです。核となるのは4ステップのループ:Read(必要なコンテキストのみ)、Constrain(明確なシステムとフォーマットルール)、Act(構造化出力、関数呼び出し、またはプレーンテキスト)、Explain(ユーザーにステップと引用を表示)。また、一般的なパターン(ルーター、抽出器、翻訳器など)、安全なリリース(テスト、監視、フォールバック)、よくある落とし穴についてもカバーしています。目標は、ユーザーが毎日依存する、目に見えない信頼性の高いAI機能を構築することです。

Groq Blogサイト内本文プロダクト内でのLLM活用:実践的フィールドガイド

OpenAIオープンセーフティモデルのデイゼロサポート

GroqCloudは、OpenAIの最新オープンソースセーフティモデルGPT-OSS-Safeguard-20Bを即日サポートし、1000 t/s超の推論速度を提供します。このモデルはセーフティ分類ワークロード向けに設計され、独自ポリシーの持ち込み、設定可能な推論努力、完全な推論トレースを備え、コストはベースモデルと同じです。

Groq Blogサイト内本文OpenAIオープンセーフティモデルのデイゼロサポート

GroqCloud でリモート MCP サポートのベータ版を開始

Groq は GroqCloud 上で MCP コネクタのベータ版を発表しました。最初に Google Workspace(Gmail、ドライブ、カレンダー)をサポートします。これらのプリビルドされた Groq ホストの MCP サーバーにより、AI エージェントは Responses API を介して Google ツールと対話でき、独自の MCP サーバーを管理する必要がありません。

Groq Blogサイト内本文GroqCloud でリモート MCP サポートのベータ版を開始

Groqが2025年Gartner® AIインフラストラクチャのクールベンダーに選出

Groqは、Gartnerの2025年AIインフラストラクチャレポートでクールベンダーに選ばれました。LPUチップの決定論的で低レイテンシな推論と線形スケーリングが評価されています。250万人以上の開発者がGroqを利用し、GPU比で最大5倍高速で低コストです。

Groq Blogサイト内本文Groqが2025年Gartner® AIインフラストラクチャのクールベンダーに選出

アメリカのAIスタックを前進させる

記事は、アメリカのAIコンピューティング、特に推論におけるリーダーシップと、その優位性を維持するための輸出政策について論じています。市場主導のエコシステムと業界連合の役割を強調し、柔軟なマルチモデルフレームワークを提案しています。

Groq Blogサイト内本文アメリカのAIスタックを前進させる

GroqCloud:需要に応える拡大

GroqCloudは、リアルタイムアプリケーションが実験から本番へ移行する需要の高まりに応えるため、AI推論インフラをグローバルに拡大しています。最近英国にEquinixと提携して新データセンターを開設し、ヨーロッパの開発者や企業に低遅延で高性能な推論を提供します。GroqCloudは現在350万人以上の開発者を擁し、本番トラフィックは増加し続けています。

Groq Blogサイト内本文GroqCloud:需要に応える拡大

LPUの内部構造:Groqの速度を解き明かす

GroqのLPUは推論専用に設計されたハードウェアであり、TruePoint数値方式、SRAMストレージ、静的スケジューリング、テンソル並列処理などを通じて、精度を犠牲にすることなく超低遅延推論を実現します。MoonshotのKimi K2はGroq上で40倍のパフォーマンスを発揮し、LPUアーキテクチャの優位性を示しています。

Groq Blogサイト内本文LPUの内部構造:Groqの速度を解き明かす

全ソース