本文にスキップ
AI News HubLIVE

GPU インフラの最新ニュース

翻訳待ち:Sakana AI Launches Fugu Max and Fugu Ultra v2 for Cheaper, Stronger Multi-Agent Orchestration

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Sakana AI has released Fugu Max and Fugu Ultra v2, 2 models built on the same learned orchestration architecture. Fugu Max routes tasks to lean open and specialized models, including NVIDIA Nemotron, at $2/$6 per 1M tokens. Fugu Ultra v2 targets peak capability, scoring 48.3 on Chartography and 74.3 on DeepSWE. The post Sakana AI Launches Fugu Max and Fugu Ultra v2 for Cheaper, Stronger Multi-Agent Orchestration appeared first on MarkTechPost.

MarkTechPostサイト内本文翻訳待ち:Sakana AI Launches Fugu Max and Fugu Ultra v2 for Cheaper, Stronger Multi-Agent Orchestration

翻訳待ち:Testing Between the Test Cases: Proving End-to-End Steering in Conditions You Never Drove

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:arXiv:2609.10951v1 Announce Type: new Abstract: AI-based automated vehicle testing is challenging because a model that passes every test condition can still fail in the real world. Formal verification offers a way to directly address this gap. On a simulated highway and an arterial road we trained two small end-to-end steering networks each in CARLA, one on clear conditions alone and one on clear, fog, night and low sun. All four models were driven against a 2.19 ft lane-departure budget. Without driving again, we used bound propagation, a formal method that reads the trained weights, to compute how far steering can drift at every disturbance strength between two captured images. One calculation covers more than a campaign could drive: on the arteri…

arXiv Roboticsサイト内本文翻訳待ち:Testing Between the Test Cases: Proving End-to-End Steering in Conditions You Never Drove

翻訳待ち:NVIDIA Details BioNeMo Inference Runtime (BioIR): 2.90x Higher Boltz-2 Folding Throughput and 58.5K Residues per GPU-Hour on 8xH100

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:NVIDIA has detailed BioNeMo Inference Runtime (BioIR), a Python library that accelerates biomolecular structure-prediction models on NVIDIA GPUs while staying in plain PyTorch. In a matched benchmark on 1,000 human dimer targets across 8xH100 GPUs, BioIR-accelerated Boltz-2 delivered 58.5K successfully folded residues per GPU-hour versus 20.2K for a torch-compiled open-source implementation, a 2.90x gain. The runtime optimizes at 3 layers: custom kernel selection, CUDA Graph capture, and Ray-based replica scaling that places 1 full model copy per GPU. BioIR already powered the AlphaFold Database expansion, generating about 31 million candidate protein complexes across 4,777 proteomes. The post NVIDIA Details BioNeMo Inference Runtime (BioIR): 2.90x…

MarkTechPostサイト内本文翻訳待ち:NVIDIA Details BioNeMo Inference Runtime (BioIR): 2.90x Higher Boltz-2 Folding Throughput and 58.5K Residues per GPU-Hour on 8xH100

翻訳待ち:Reduce inference cold starts on Amazon SageMaker HyperPod with model caching

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Amazon SageMaker HyperPod now supports model caching for inference, which pre-loads model weights and container images onto cluster nodes so pods read from local NVMe storage instead of downloading over the network. Learn how model caching cuts cold starts from tens of minutes to seconds, how it works, and how to enable it.

AWS Machine Learning Blogサイト内本文翻訳待ち:Reduce inference cold starts on Amazon SageMaker HyperPod with model caching

翻訳待ち:Google to Invest $15B in Finland’s AI Infrastructure

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:The tech giant simultaneously revealed a nuclear power contract with Finnish operator Fortum, its first outside of the U.S.

AI Businessサイト内本文翻訳待ち:Google to Invest $15B in Finland’s AI Infrastructure

翻訳待ち:Red Hat AI 3.5 tackles the GPU queue that can stall AI pilots

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Red Hat released Red Hat AI 3.5 this week, a move designed to let software engineering teams run AI with The post Red Hat AI 3.5 tackles the GPU queue that can stall AI pilots appeared first on The New Stack.

The New Stack AIサイト内本文翻訳待ち:Red Hat AI 3.5 tackles the GPU queue that can stall AI pilots

翻訳待ち:Skild AI Taps NVIDIA Physical AI to Teach Robots New Tasks From a Single Video

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Manufacturing floors, warehouses and production lines rarely stay fixed — tasks change, layouts shift and new products arrive, and most robots can’t keep up without significant reprogramming. Skild AI’s new S1 robot foundation model helps address this, designed to learn previously unseen, long-horizon tasks from a single video demonstration. The model, launched last week, uses […]

NVIDIA Blogサイト内本文翻訳待ち:Skild AI Taps NVIDIA Physical AI to Teach Robots New Tasks From a Single Video

翻訳待ち:Physical AI Takes the Wheel: How the World’s Robotaxi Leaders Are Building With NVIDIA Technologies

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:The global robotaxi market — physical AI’s first commercial breakthrough — is projected to reach $400 billion by 2035, with over 6 million commercial vehicles in operation as driverless fleets are already moving people through some of the world’s busiest and most complex streets. Deploying a driverless vehicle is one challenge. Scaling a fleet is […]

NVIDIA Blogサイト内本文翻訳待ち:Physical AI Takes the Wheel: How the World’s Robotaxi Leaders Are Building With NVIDIA Technologies

翻訳待ち:Universal Music is launching an AI music platform with ElevenLabs

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Universal Music Group is launching a new AI-powered platform that will allow users to draw from its catalog of licensed music to create song remixes, mashups, and new takes on tracks, according to an announcement on Thursday. The record label is developing the platform through a multi-year licensing agreement with ElevenLabs, a company that specializes in AI voice and music generation. Artists can choose whether to participate in UMG and ElevenLabs' upcoming platform, which marks yet another AI deal for the record label. UMG is currently developing an AI music platform with Udio and has struck AI licensing deals with Spotify, Nvidia, and Kl … Read the full story at The Verge.

The Verge AIサイト内本文翻訳待ち:Universal Music is launching an AI music platform with ElevenLabs

翻訳待ち:d-Matrix Adopts NVIDIA NVLink Fusion for Rack-Scale XPU Deployment

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:AI inference chipmaker d-Matrix today announced it will use NVIDIA NVLink Fusion to connect its next-generation Raptor XPUs to NVIDIA’s AI infrastructure platform — joining a growing roster of ecosystem partners. By connecting Raptor to NVIDIA NVLink scale-up and Spectrum-X scale-out networking, the NVIDIA MGX rack architecture and the broader NVIDIA AI platform, NVLink Fusion […]

NVIDIA Blogサイト内本文翻訳待ち:d-Matrix Adopts NVIDIA NVLink Fusion for Rack-Scale XPU Deployment

翻訳待ち:Boots on the Ground: ‘WARDOGS’ Goes All Out on GeForce NOW at Early-Access Launch

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Gear up: The latest PC games and major updates are ready to play on GeForce NOW this week. WARDOGS drops onto the cloud at early-access launch, alongside the Valheim 1.0 Deep North update and Bus Simulator 27 — part of nine new titles joining the cloud. The newest PC releases can demand serious hardware, storage […]

NVIDIA Blogサイト内本文翻訳待ち:Boots on the Ground: ‘WARDOGS’ Goes All Out on GeForce NOW at Early-Access Launch

翻訳待ち:“AI factories are among the most complex systems ever built”: Nvidia and Palantir turn Nvidia’s supply chain into a proving ground for sovereign AI

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Nvidia and Palantir announced on Thursday that they’re working together to bring “sovereign AI to critical supply chains,” kicking off The post “AI factories are among the most complex systems ever built”: Nvidia and Palantir turn Nvidia’s supply chain into a proving ground for sovereign AI appeared first on The New Stack.

The New Stack AIサイト内本文翻訳待ち:“AI factories are among the most complex systems ever built”: Nvidia and Palantir turn Nvidia’s supply chain into a proving ground for sovereign AI

翻訳待ち:Introducing preemptible compute: the same compute, half the price

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Together GPU Clusters now supports preemptible compute: the same GPU capacity at a flat 50% of the on-demand rate, with a five-minute drain window.

Together AI Blogサイト内本文翻訳待ち:Introducing preemptible compute: the same compute, half the price

翻訳待ち:To Infinity and Beyond: ThunderKittens Now on NVIDIA Vera Rubin NVL72!

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:We ported ThunderKittens to NVIDIA's Vera Rubin NVL72 and rebuilt our NVFP4 GEMM around the new hardware, taking it from 42% of roofline to over 22 PFLOPS — competitive with cuBLAS and CuTe DSL. Here is what changed in the ISA and how we used it.

Together AI Blogサイト内本文翻訳待ち:To Infinity and Beyond: ThunderKittens Now on NVIDIA Vera Rubin NVL72!

翻訳待ち:Deploying Qwen3.8-2.4T-A95B on Amazon SageMaker HyperPod with vLLM

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Learn how to deploy Qwen3.8-2.4T-A95B, a 2.4-trillion-parameter open-weight model, on Amazon SageMaker HyperPod with vLLM. This walkthrough covers cluster provisioning, NVFP4 quantization, and an OpenAI-compatible endpoint with built-in reasoning, tool calling, and native MTP speculative decoding.

AWS Machine Learning Blogサイト内本文翻訳待ち:Deploying Qwen3.8-2.4T-A95B on Amazon SageMaker HyperPod with vLLM

翻訳待ち:Qualcomm Forges AI Chip Deal with Amazon

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:The chipmaker is competing with Nvidia, the world’s dominant producer of AI processors.

AI Businessサイト内本文翻訳待ち:Qualcomm Forges AI Chip Deal with Amazon

翻訳待ち:NVIDIA Brings Real-Time AI to Broadcast, Sports and Global Streaming at IBC

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:At the IBC conference, running Sept. 11-14 in Amsterdam, the creative, technology and business communities are coming together to turn ideas into action and discuss innovations across the media and entertainment industries. More than 44,000 attendees from 170+ countries are gathering to explore 1,300+ exhibitions in 14+ halls and outdoor spaces, with over 600 speakers […]

NVIDIA Blogサイト内本文翻訳待ち:NVIDIA Brings Real-Time AI to Broadcast, Sports and Global Streaming at IBC

翻訳待ち:Simplify and support your TorchServe workloads using Ray Serve Deep Learning Containers

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:TorchServe is no longer maintained, leaving teams to own the entire GPU inference stack. The AWS Ray Serve Deep Learning Container is a supported, pre-tested container with the framework, GPU drivers, and serving layer already assembled. This post walks through deploying a vision-language model on Amazon EKS using the Ray Serve DLC on a single GPU node.

AWS Machine Learning Blogサイト内本文翻訳待ち:Simplify and support your TorchServe workloads using Ray Serve Deep Learning Containers

翻訳待ち:7 Approaches to Efficient LLM Training on Limited Hardware

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Learn seven engineering techniques to train large language models on consumer GPUs without running out of memory.

KDnuggetsサイト内本文翻訳待ち:7 Approaches to Efficient LLM Training on Limited Hardware

Perplexity、GPU埋め込みスタックを解説:Ivy、Tulip、ROSEがpplx-embedを支える仕組み

Perplexityのエンジニアリングチームは、埋め込みモデルpplx-embedのGPU推論基盤を解説する記事を公開。検索品質はモデル品質と運用コストに左右されるとし、LLM用のprefill/decodeカーネルを再利用する設計を紹介。Rust製ゲートウェイIvy、gRPCサーバTulip、Python製エンジンROSEの3層構成や、CUDA Graphによる起動オーバーヘッド削減、LazyTensorによる非同期処理などの工夫が述べられている。

MarkTechPostサイト内本文Perplexity、GPU埋め込みスタックを解説:Ivy、Tulip、ROSEがpplx-embedを支える仕組み

Nous Research、Hermes Desktopにワンクリックのローカルモデルセットアップ機能を追加

Nous Researchは、Hermes Desktopでのローカルオープンウェイトモデルのセットアップをワンクリックに簡素化しました。アプリがハードウェアを読み取り、GPUに合うモデルを選択し、重みをダウンロードしてllama.cppを自動構成します。アカウント不要で利用可能。量子化は4ビット下限、推奨モデルは64K以上のコンテキストウィンドウを保証し、各モデルのメモリ適合状況を緑・黄・赤で表示します。

MarkTechPostサイト内本文Nous Research、Hermes Desktopにワンクリックのローカルモデルセットアップ機能を追加

MicrosoftのProject Zenithは開発者向けの「気が散らないWindowsエクスペリエンス」

マイクロソフトは開発者向けに最適化したWindowsを「Project Zenith」として正式に発表した。64GB以上のユニファイドメモリを備えたデバイス向けで、事前構成済みのWindows環境と厳選された開発者ツールを提供。ローカルで300億パラメータ以上のモデルを従量課金なしで実行できる。AMDがIFAで最初のProject ZenithミニPCを発表し、今後も異なるシリコンを搭載した製品が登場する予定だ。

The Verge AIサイト内本文MicrosoftのProject Zenithは開発者向けの「気が散らないWindowsエクスペリエンス」

Hugging Face買収でNVIDIA、AI競争に大勝利

AI競争ではAnthropic、OpenAI、Googleなどのモデル開発者に注目が集まりがちだが、いま主役となっているのはNVIDIAだ。Hugging Faceの買収によって、その地位をさらに強固にしている。

AI Businessサイト内本文Hugging Face買収でNVIDIA、AI競争に大勝利

「Hugging Faceはオープンプラットフォームのまま」、Nvidiaが129億ドルで“AIのGitHub”買収

Nvidiaは、AIモデルのオープンプラットフォーム「Hugging Face」を約129億ドルで買収することで合意した。買収後もプラットフォームの開放性とハードウェア中立性を維持し、Nvidia製コンピューティングの利用は必須にしないと同社は約束。取引は規制当局の承認などを条件に2027年前半の完了を見込む。

The New Stack AIサイト内本文「Hugging Faceはオープンプラットフォームのまま」、Nvidiaが129億ドルで“AIのGitHub”買収

Nvidia PAIR、アイドル状態のMacとPCをAIエージェント向けに活用

Nvidiaが発表したオープンソースのソフトウェアルーター「PAIR」は、自宅のネットワークにあるアイドル状態のMacやPCを利用してAIモデルを実行し、サブエージェントによるエージェントワークフローを高速化する。Ollama/LM Studioに対応し、ベータ版として提供中。

The New Stack AIサイト内本文Nvidia PAIR、アイドル状態のMacとPCをAIエージェント向けに活用

Nvidia、アイドル状態のPCをつないで“個人AIデータセンター”にする無償ツール「PAIR」を発表

エヌビディアは、家庭ネットワーク上のアイドル状態にある互換PCを自動検出して接続し、ローカルAI推論やエージェント型ワークフローに活用できる無償のオープンソースソフトウェア「PAIR」を発表した。対応するのは GeForce RTX 20シリーズ以降、RTX Pro GPU、DGX Spark、Apple M4以降。ベータ版は本日からWindows、Linux、macOSで利用できる。

The Verge AIサイト内本文Nvidia、アイドル状態のPCをつないで“個人AIデータセンター”にする無償ツール「PAIR」を発表

『NBA 2K27』がNVIDIA DLSS 5を引っ提げ、GeForce NOWに登場する今月の新作26タイトルを牽引

GeForce NOWは9月に26のゲームを追加し、目玉はNVIDIA DLSS 5の3Dガイド型ニューラルレンダリングを採用した『NBA 2K27』。『The Blood of Dawnwalker』や『鬼武者: Way of the Sword』も発売日からクラウドでプレイでき、UltimateメンバーはRTX 5080級のクラウド性能を利用できます。

NVIDIA Blogサイト内本文『NBA 2K27』がNVIDIA DLSS 5を引っ提げ、GeForce NOWに登場する今月の新作26タイトルを牽引

NVIDIA、Hugging Faceの買収を発表

NVIDIAは、Hugging Faceを129億3030万ドルで買収することに合意した。プラットフォームのオープン性を維持しながら、インフラとAIエコシステムへのアクセスを拡大する。

NVIDIA Blogサイト内本文NVIDIA、Hugging Faceの買収を発表

ZimaBlue:スケーラブルなビデオ事前学習による汎用世界行動モデルの進化

本論文では、大規模ビデオから汎用世界行動モデル(WAM)を学習するためのスケーラブルなフレームワークであるZimaBlueを紹介します。三段階のトレーニングカリキュラム(大規模な人間・ロボットの自己中心ビデオでの因果的身体化事前学習、統一行動表現を用いたビデオ-行動中間学習、ロボット固有の特化)を採用しています。非同期のSlow-Fastデュアルアーキテクチャにより、NVIDIA RTX 4090上で30Hzのリアルタイム行動予測を実現します。実ロボットのゼロショット評価では、ターゲットロボットデータのみから120,000時間以上の身体化ビデオにスケーリングすることで成功率が36.1%から77.8%に向上しました。特に未見のタスクで顕著な向上が見られます。

arXiv Computer Visionサイト内本文ZimaBlue:スケーラブルなビデオ事前学習による汎用世界行動モデルの進化

CUDA-Harness:自然言語からのエージェント型CUDAカーネル生成と最適化

CUDA-Harnessは、自然言語から高性能CUDAカーネルを生成・最適化するための新しいフレームワークです。高レベルの意味理解と低レベルのカーネル生成を結びつける中間構造化生成を導入し、総合ベース検証により報酬ハッキングを軽減し、フィードバック適応進化により正確性を優先しながら性能を最適化します。実験により、その有効性と、LLM・ハードウェア・C-to-CUDA移行への汎用性が実証されています。

arXiv Computational Linguisticsサイト内本文CUDA-Harness:自然言語からのエージェント型CUDAカーネル生成と最適化

Modal上でフルスタック生成AIを実行するBotikaの取り組み

ファッションECブランド向けのビジュアル生成AIを手がけるBotikaは、社内で開発した基盤モデルと、100TBの画像データを処理するパイプライン、約15モデルの本番推論をすべてModal上で運用。CEOのEran Dagan氏は、Modalがなければ研究スピードやインフラ負担の面で現在の体制は実現できなかったと語る。

Modal Blogサイト内本文Modal上でフルスタック生成AIを実行するBotikaの取り組み

NVIDIAとCrowdStrike、エージェント型サイバーセキュリティの最前線を強化

CrowdStrikeのFal.Con 2026で、NVIDIA CEOのジェンスン・フアン氏とCrowdStrike CEOのジョージ・カーツ氏が、SafeMindを発表。NVIDIA NemotronモデルとCrowdStrikeの脅威データを組み合わせ、攻撃と防御のAIを継続的に共進化させることで、自動化された攻撃に対抗する。

NVIDIA Blogサイト内本文NVIDIAとCrowdStrike、エージェント型サイバーセキュリティの最前線を強化

Show HN: オープンソースのK8sネイティブAIプラットフォーム、分散マルチモデル推論用

shaideは、自社のKubernetesクラスター上で大規模にAIモデルを提供するセルフホスト型AIプラットフォームです。単一コマンドでインストールでき、ネットワーク境界内に完全に収まり、完全にエアギャップされた環境もサポートします。このプラットフォームはPulumiを使用してインフラをコードとして管理し、OpenAI互換APIを提供し、エージェント群向けに最適化されています。

Hacker News AIサイト内本文Show HN: オープンソースのK8sネイティブAIプラットフォーム、分散マルチモデル推論用

NVIDIAの物議を醸すDLSS 5が9月3日に登場、ハイエンドGPU性能が必要に

NVIDIAは今週、DLSS 5を正式にリリースするが、対応ゲームは『NBA 2K27』のみで、中級GPUのRTX 5060でも6倍フレーム生成を有効にしなければ1080pで快適に動作しないほどの高い性能が求められる。

The Verge AIサイト内本文NVIDIAの物議を醸すDLSS 5が9月3日に登場、ハイエンドGPU性能が必要に

Nemotron 3 Ultra 解説:NVIDIA の 550B ハイブリッド Mamba-MoE モデル

本記事では、NVIDIA が 2026 年 6 月 4 日に公開した Nemotron 3 Ultra について詳説します。これは 5500 億パラメータのハイブリッド Mamba-アテンション MoE モデルで、トークンあたり約 550 億パラメータのみを活性化(スパーシティ約 10%)し、Mamba-2 状態空間層と Transformer アテンション層を組み合わせて長期的なエージェントタスク向けに設計されています。Nemotron 3 ファミリー(Nano、Super、Ultra)の位置づけ、ハイブリッドアーキテクチャの動機、トレーニング詳細、および OpenRouter、NVIDIA NIM、またはセルフホストの vLLM による呼び出し方法を説明します。

Hacker News AIサイト内本文Nemotron 3 Ultra 解説:NVIDIA の 550B ハイブリッド Mamba-MoE モデル

翻訳待ち:Private AI search across your work

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Private AI search across your work. Ask across Google Drive, Dropbox, and Slack. Get source-backed answers without a permanent index. Choose German hosting or your own infrastructure. Create free account Create your acc…

Hacker News AIサイト内本文翻訳待ち:Private AI search across your work

翻訳待ち:The AI moat isn't GPUs, it's the advanced packaging they require

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:The Hidden Bottleneck: Why Advanced Packaging is the Next AI Moat Executive Summary The market's obsession with GPU design is noise. The true bottleneck—and strategic moat—in the AI hardware race is not the silicon itse…

Hacker News AIサイト内本文翻訳待ち:The AI moat isn't GPUs, it's the advanced packaging they require

翻訳待ち:Nvidia and Semiconductor Vendor Expand Partnership

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:The move shows how the AI hardware giant is aiming to maintain its dominance, even through third parties.

AI Businessサイト内本文翻訳待ち:Nvidia and Semiconductor Vendor Expand Partnership

翻訳待ち:AWS recognized as a Leader in The Forrester Wave: AI Infrastructure Solutions, Q4 2025

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:We're excited to share that AWS has been recognized as a Leader in The Forrester Wave: AI Infrastructure Solutions, Q4 2025. In this evaluation of 13 providers, AWS received the highest score in the Strategy category.

AWS Machine Learning Blogサイト内本文翻訳待ち:AWS recognized as a Leader in The Forrester Wave: AI Infrastructure Solutions, Q4 2025

翻訳待ち:Trump says datacenter opponents ‘want to end up being backwards and poor’

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Three-quarters of Americans said in a recent poll that they oppose datacenters being built next to their homes Donald Trump has criticized communities pushing back against datacenter projects across the US amid a growing backlash, warning those that reject them risk becoming “backwards and poor”. As controversy surrounding local datacenter plans continues to swirl around election campaigns nationwide ahead of November’s midterm elections, the US president declared Americans “will only have yourselves to blame” if they are canceled. Continue reading...

The Guardian AIサイト内本文翻訳待ち:Trump says datacenter opponents ‘want to end up being backwards and poor’

翻訳待ち:Apple Is Suddenly an AI Infra Stock as OpenAI Buys 10k+ Macs

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Investing Apple Is Suddenly an AI Infrastructure Stock as OpenAI Buys Macs by the Tens of Thousands OpenAI has been quietly buying Apple hardware by the tens of thousands, and it has nothing to do with iPhones or consum…

Hacker News AIサイト内本文翻訳待ち:Apple Is Suddenly an AI Infra Stock as OpenAI Buys 10k+ Macs

翻訳待ち:Speed Up LLM Inference with DSpark Speculative Decoding

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Learn how DSpark speculative decoding can improve local LLM generation speed using the same GPU, with Qwen3-8B, llama.cpp, and CUDA.

KDnuggetsサイト内本文翻訳待ち:Speed Up LLM Inference with DSpark Speculative Decoding

翻訳待ち:New York Governor Kathy Hochul thinks AI should be ‘less evil’

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Today, I’m talking with New York Governor Kathy Hochul, and I’ll just warn you — this episode moves really fast. It’s an election year, after all, with a shocking amount of tech policy at stake, and Governor Hochul has taken strong positions on almost every major tech issue there is. For example, Meta just reached a settlement with dozens of states, including New York, which will restrict how teens use platforms like Instagram in very specific ways. Governor Hochul is a strong supporter of those restrictions and more, as you’ll hear. But those come with a cost — widespread age verification means adults will also have to show ID to use the internet, which will essentially make it impossible to be anonymous online. Verge subscribers, don’t forget you…

The Verge AIサイト内本文翻訳待ち:New York Governor Kathy Hochul thinks AI should be ‘less evil’

翻訳待ち:Beyond Data Scaling: Representation-Centric Continued Pre-training for Vision-Language-Action Models

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:arXiv:2608.27550v1 Announce Type: new Abstract: Scaling robot data is crucial for building generalist Vision-Language-Action (VLA) models, yet robot trajectories are harder to scale than web-scale image-text data because embodied collection is costly and sparsely covers the physical world. This makes representation quality a central bottleneck: under a fixed robot-data budget, continued pre-training must turn limited trajectories into transferable visual-action knowledge rather than merely fit actions. We propose VLAct, a VLA-oriented VLM backbone trained on broad, heterogeneous, multi-embodiment robot data before task-specific fine-tuning. VLAct preserves the broad VLM prior and encourages shared action semantics across embodiments through VLM-prio…

arXiv Roboticsサイト内本文翻訳待ち:Beyond Data Scaling: Representation-Centric Continued Pre-training for Vision-Language-Action Models

翻訳待ち:ABCD: Alpha-Composited Block Coordinate Descent: Constant-VRAM Training for Large Radiance Fields

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:arXiv:2608.27735v1 Announce Type: new Abstract: We present ABCD (Alpha-Composited Block Coordinate Descent), an out-of-core training framework for alpha-composited radiance fields, instantiated here for 3D Gaussian Splatting. Our method reformulates training as block coordinate descent over spatial partitions: only one block of parameters is active at a time, while all others are frozen. By exploiting the associativity of alpha blending, these inactive regions can be pre-rendered and collapsed into foreground and background RGBA images. As a result, for fixed partition size and image resolution, peak VRAM becomes O(1) with respect to total scene extent, rather than growing with full scene size. This enables GPUs with limited memory to train scenes t…

arXiv Computer Visionサイト内本文翻訳待ち:ABCD: Alpha-Composited Block Coordinate Descent: Constant-VRAM Training for Large Radiance Fields

翻訳待ち:Quanta Perception as Probabilistic Events

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:arXiv:2608.27584v1 Announce Type: new Abstract: Autonomous systems rely on extracting information from light, yet remain brittle in extreme environments, from nighttime navigation to high-speed robotics. Conventional sensors aggregate photons over fixed exposures, imposing trade-offs between sensitivity, dynamic range, and temporal resolution that degrade perception when photons are scarce or dynamics are rapid. Quanta sensors detect individual photons, but their streams exceed real-time compute and latency budgets by orders of magnitude. Here we introduce $\textit{probabilistic events}$, a computational primitive for real-time quanta perception from individual photon detections. By computing the posterior over the time since the last intensity chan…

arXiv Computer Visionサイト内本文翻訳待ち:Quanta Perception as Probabilistic Events

翻訳待ち:DAMP: Decay-Aware Mixed-Precision Recurrent-State Quantization

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:arXiv:2608.27513v1 Announce Type: new Abstract: Softmax attention stores key and value vectors for every preceding token, causing inference memory to grow with sequence length. Recent language models incorporating Gated DeltaNet (GDN) or Kimi Delta Attention (KDA) reduce this cost by replacing the KV cache in most layers with fixed-size recurrent states. However, these recurrent states are commonly stored in FP32 and consume substantial GPU memory; their updates are memory-bandwidth bound and contribute significantly to decoding latency. To our knowledge, we are the first to study post-training quantization of recurrent states in GDN and KDA based language models. We find that uniform quantization provides a poor accuracy--storage trade-off: INT8 an…

arXiv Machine Learningサイト内本文翻訳待ち:DAMP: Decay-Aware Mixed-Precision Recurrent-State Quantization

翻訳待ち:AI’s worst disasters will arrive unannounced | Letters

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Dr Simon Nieder on the global efforts needed to curb the threats posed by AI. Plus letters from Dr Anthony Harris and David Kyler Timothy Garton Ash is right that serious artificial intelligence risks demand international action, but “AI Hiroshima” may be the wrong picture (Would even an AI disaster on the scale of Hiroshima be enough to make humankind protect itself? I fear not, 22 August). Hiroshima was not technology going rogue. Human beings designed the bomb, authorised its use and dropped it. The technology worked much as intended. AI may one day behave in ways we cannot control, but many of the gravest harms could happen without any dramatic moment when “the AI takes over”. AI might help design a pathogen, find a vulnerability in critical inf…

The Guardian AIサイト内本文翻訳待ち:AI’s worst disasters will arrive unannounced | Letters

翻訳待ち:Show HN: FinBridge – Korean stock market data for AI agents (MCP)

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:Your agent reads EDGAR. It cannot read DART. Your model knows NVIDIA. It guesses about SK Hynix. Ask it about both. “Find KOSDAQ names above RS 90 that pass the trend template.” Ask Claude and it screens every Korean li…

Hacker News AIサイト内本文翻訳待ち:Show HN: FinBridge – Korean stock market data for AI agents (MCP)

翻訳待ち:The Sequence Radar-Issue #923: Last Week in AI: AI’s Industrial Turn

AI サービスが一時的に利用できないため、復旧後に翻訳を補完します。ソース概要:NVIDIA, Anthropic, NScale, and a16z show how the AI race is moving from models to infrastructure.

TheSequenceサイト内本文翻訳待ち:The Sequence Radar-Issue #923: Last Week in AI: AI’s Industrial Turn

その他の成長タグ

GPU インフラ AI News | AI News Hub