AI News HubLIVE

來源分布

  • SiliconANGLE AI12
  • Hacker News AI9
  • NVIDIA Blog8
  • MarkTechPost5
  • AI Business3
  • arXiv AI2
  • Analytics Vidhya1
  • arXiv Machine Learning1

主題分布

  • 芯片49
  • Agent26
  • 研究11
  • 模型9
  • 創業融資8
  • 機械人2
  • 工具1

日期線

  • 2026-08-259
  • 2026-08-187
  • 2026-08-266
  • 2026-08-245
  • 2026-08-174
  • 2026-08-194
  • 2026-08-204
  • 2026-08-224

最新動態

待翻譯:Nvidia Jetson Orin-guided Russian AI drone killed three civilians in Ukraine

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:5 Join the conversation Follow us Add us as a preferred source on Google A Russian Molniya drone carrying an Nvidia Jetson Orin module crashed and killed three civilians at a gas station in Zaporizhzhia last month after…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • 5 Join the conversation Follow us Add us as a preferred source on Google A Russian Molniya drone carrying an Nvidia Jetson Orin module crashed and killed three civilians at a gas…
站內正文

待翻譯:Perplexity's Portable Computer tackling local AI services market

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Now Perplexity is trying to get into the local AI action Amid talk of an Nvidia deal, the AI search biz is looking beyond the cloud Thomas Claburn Thomas Claburn AI AND SOFTWARE REPORTER Published wed 26 Aug 2026 // 00:…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Now Perplexity is trying to get into the local AI action Amid talk of an Nvidia deal, the AI search biz is looking beyond the cloud Thomas Claburn Thomas Claburn AI AND SOFTWARE R…
站內正文

待翻譯:Perplexity AI launches Portable Computer on-device AI agent

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Perplexity AI Inc. today introduced Portable Computer, an artificial intelligence agent designed to run on desktops equipped with Nvidia Corp. silicon. The launch follows a report that Nvidia is weighing an investment in the startup that could value it at over $30 billion. Furthermore, Nvidia has reportedly floated the idea of licensing Perplexity’s technology and […] The post Perplexity AI launches Portable Computer on-device AI agent appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Perplexity AI Inc. today introduced Portable Computer, an artificial intelligence agent designed to run on desktops equipped with Nvidia Corp. silicon. The launch follows a report…
站內正文

待翻譯:Perplexity Ships Portable Computer on NVIDIA DGX Spark: Local Harness, OS-Enforced Sandbox, and Zero Per-Token Cost for Local Steps

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Perplexity releases Portable Computer, packaging local models, harness, sandbox, and connectors into one system running on NVIDIA DGX Spark. The post Perplexity Ships Portable Computer on NVIDIA DGX Spark: Local Harness, OS-Enforced Sandbox, and Zero Per-Token Cost for Local Steps appeared first on MarkTechPost.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Perplexity releases Portable Computer, packaging local models, harness, sandbox, and connectors into one system running on NVIDIA DGX Spark. The post Perplexity Ships Portable Com…
站內正文

待翻譯:Taiwan charges nine people for smuggling ‘high-end’ AI servers to China

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Among those charged are two Super Micro employees and one from Nvidia, marking another flashpoint in US-China AI rivalry Taiwanese prosecutors charged nine people Monday, including one from Nvidia and two from Super Micro, for illegally exporting “high-end AI servers” to mainland China, adding another wave of turbulence in the AI ​​rivalry between China and the United States. Prosecutors said the servers involved were graphics processing units known as “B300,” which have been banned from sale to China. Continue reading...

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Among those charged are two Super Micro employees and one from Nvidia, marking another flashpoint in US-China AI rivalry Taiwanese prosecutors charged nine people Monday, includin…
站內正文

待翻譯:Nvidia's Groq 3 LPX accelerator enters full production at 3,500 tokens/SEC

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production as it strives to maintain its dominance in the world of AI compute. The new chip, announ…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production as it strives to maintain its dominance in the w…
站內正文

待翻譯:Perplexity’s Computer agent can now run locally — if you can afford it

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Perplexity, in partnership with Nvidia, has taken Computer, its agentic AI assistant, and brought it to the desktop in the The post Perplexity’s Computer agent can now run locally — if you can afford it appeared first on The New Stack.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Perplexity, in partnership with Nvidia, has taken Computer, its agentic AI assistant, and brought it to the desktop in the The post Perplexity’s Computer agent can now run locally…
站內正文

待翻譯:Leading Publishers Bring Blockbuster PC Games and Technology to NVIDIA RTX Spark

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:NVIDIA is bringing the next wave of RTX gaming to the Gamescom conference running this week in Cologne, Germany, with support for new games, anti-cheat technologies and increased visual quality. Electronic Arts, Embark and Ubisoft are among the latest game publishers and developers bringing their blockbuster titles to NVIDIA RTX Spark ahead of its launch […]

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • NVIDIA is bringing the next wave of RTX gaming to the Gamescom conference running this week in Cologne, Germany, with support for new games, anti-cheat technologies and increased…
站內正文

待翻譯:Nvidia doubles compute for entry-level edge robotics with Jetson Orin Nano 2

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Nvidia Corp. today announced the release of Jetson Orin Nano 2, a robotics computer “brain” for running artificial intelligence and frontier-level models at the edge. In the past months, foundational AI models have grown smaller and more efficient, adding numerous capabilities alongside language understanding, computer vision and audio processing. As more AI models compress in […] The post Nvidia doubles compute for entry-level edge robotics with Jetson Orin Nano 2 appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Nvidia Corp. today announced the release of Jetson Orin Nano 2, a robotics computer “brain” for running artificial intelligence and frontier-level models at the edge. In the past…
站內正文

待翻譯:AI factories enter the execution era as Cisco and NVIDIA push rack-scale systems into production

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The artificial intelligence infrastructure market is crossing an important threshold. The conversation is shifting from acquiring graphics processing units to building complete AI factories that can generate tokens reliably, efficiently and at scale. This is the next bottleneck. GPUs may be the engine, but an AI factory is a system. Compute, networking, storage, cooling, software […] The post AI factories enter the execution era as Cisco and NVIDIA push rack-scale systems into production appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The artificial intelligence infrastructure market is crossing an important threshold. The conversation is shifting from acquiring graphics processing units to building complete AI…
站內正文

待翻譯:A Go dependency wrote AGENTS.md mid-build and got Codex to hide the change

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:In a proof of concept published in mid-2026, NVIDIA's AI Red Team planted a malicious dependency in a Go project, let it execute during a normal build, and watched it write a new AGENTS.md file into the project director…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • In a proof of concept published in mid-2026, NVIDIA's AI Red Team planted a malicious dependency in a Go project, let it execute during a normal build, and watched it write a new…
站內正文

待翻譯:Nvidia and Cisco push the enterprise AI factory into the rack-scale era

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The Cisco Secure AI Factory with Nvidia has been extended into the rack-scale era, offering enterprises greater full-stack operational capabilities as a result. Designed to give customers a framework for deploying artificial intelligence across their entire infrastructure, Secure AI Factory with Nvidia from Cisco Systems Inc. now integrates rack-to-fabric liquid cooling supporting systems beyond 200 […] The post Nvidia and Cisco push the enterprise AI factory into the rack-scale era appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The Cisco Secure AI Factory with Nvidia has been extended into the rack-scale era, offering enterprises greater full-stack operational capabilities as a result. Designed to give c…
站內正文

待翻譯:Cisco and Nvidia take AI factories from rack to runtime

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:AI factories are moving from ambitious plans toward production, but the path from graphics processing unit acquisition to usable systems remains a race against time. Neoclouds already have customers waiting for capacity, enterprises are looking to bring inference workloads closer to home and sovereign AI programs are being built now. Those distinct buyer motions are […] The post Cisco and Nvidia take AI factories from rack to runtime appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • AI factories are moving from ambitious plans toward production, but the path from graphics processing unit acquisition to usable systems remains a race against time. Neoclouds alr…
站內正文

待翻譯:Robotics AI startup Generalist reportedly raises $200M

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Generalist AI Inc., a startup that develops artificial intelligence software for robots, has reportedly raised $200 million in funding. Axios today cited a source as saying that 8VC led the investment. It was reportedly joined by a number of unnamed existing investors. Generalist’s previous $400 million round in June included the participation of Nvidia Corp., […] The post Robotics AI startup Generalist reportedly raises $200M appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Generalist AI Inc., a startup that develops artificial intelligence software for robots, has reportedly raised $200 million in funding. Axios today cited a source as saying that 8…
站內正文

待翻譯:Nvidia reportedly eyes another investment in Perplexity AI at a $30B valuation

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Nvidia Corp. is reportedly considering making another investment in the artificial intelligence search startup Perplexity AI Inc. A report by The Information says the chipmaker is holding talks with Perplexity over an investment that could push the startup’s valuation to more than $30 billion. That would represent a jump of more than 50% from the […] The post Nvidia reportedly eyes another investment in Perplexity AI at a $30B valuation appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Nvidia Corp. is reportedly considering making another investment in the artificial intelligence search startup Perplexity AI Inc. A report by The Information says the chipmaker is…
站內正文

待翻譯:How XPUs Meet a World-Class AI Factory

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:To generate intelligence at scale, AI factories run continuously, and their economics are defined by delivered output: tokens per second, tokens per watt, cost per token, utilization and uptime. That requires AI infrastructure designed and built as a full factory, not a collection of individual accelerators. Hyperscalers and AI-native companies building custom XPUs must consider […]

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • To generate intelligence at scale, AI factories run continuously, and their economics are defined by delivered output: tokens per second, tokens per watt, cost per token, utilizat…
站內正文

待翻譯:With Groq 3 LPX in Full Production, NVIDIA Extends Vera Rubin Inference for Agents

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The next era of AI inference won’t be defined by a single breakthrough chip, network or system. It’ll be defined by how every layer of the AI factory works together. That’s why NVIDIA is extending Vera Rubin NVL72 with fast token generation for agentic systems. Announced today, the NVIDIA Vera Rubin rack-scale system NVIDIA Groq […]

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The next era of AI inference won’t be defined by a single breakthrough chip, network or system. It’ll be defined by how every layer of the AI factory works together. That’s why NV…
站內正文

待翻譯:Up to 30x More Work Per Watt: NVIDIA Vera Rubin NVL72 Sets a New Efficiency Standard for AI Agents

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:According to OpenRouter data, agentic AI workloads consume 15x more tokens than a simple chat request. Why? Consider what happens when an AI agent researches a company for an investment decision. The agent queries financial databases, searches news and filings, invokes a sub-agent to run peer comparisons and model valuations, then synthesizes everything into a […]

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • According to OpenRouter data, agentic AI workloads consume 15x more tokens than a simple chat request. Why? Consider what happens when an AI agent researches a company for an inve…
站內正文

待翻譯:Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself as an inference cloud after licensing its LPU technology to NVIDIA. This comparison checks each provider's live rate card, Q2 2026 financials, active and contracted gigawatts, anchor contracts, and SemiAnalysis ClusterMAX tier. Nebius posts the lowest H100 rate and the only published B300 price, Lambda has the cheapest B200, Crusoe is the only one with AMD on its card, and CoreWeave commands a 10–15% premium as the sole Platinum-rated provider. Figures verified August 21, 2026. The post Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power appeared first on MarkTechPost.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself…
站內正文

待翻譯:BF1: A Causal Dyadic Sparse-Attention Retrofit for Efficient Long-Context Transformers

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2608.20427v1 Announce Type: new Abstract: Dense causal attention remains expensive at long context even when implemented with highly optimized exact kernels. We study BF1, a deterministic block-aligned dyadic sparse-attention route that combines a small exact local neighborhood, a global first block, and logarithmically spaced historical blocks. The route is related to prior log-sparse and dilated attention patterns; our contribution is a correctness-gated pretrained-model retrofit, a matched topology-control study, and a systems characterization that connects per-layer sparsity to whole-model latency. For fixed block width, every converted layer uses O(n log n) selected token interactions and has O(log n) graph communication depth. On an NVIDIA RTX PRO 6000 Blackwell GPU, an optimized BF16 implementation crosses dense attention between 2K and 4K tokens and reaches a 10.91x per-layer prefill speedup at 32K. Retrofitting eight of 28 Qwen3-0.6B attention layers lowers warm whole-model time to first token by 7.7%, 11.3%, and 15.3% at 8K, 16K, and 32K, respectively, while the remaining dense layers keep the complete model asymptotically quadratic. Under a matched 1,000-step, 16.384M-token adaptation protocol, BF1 ranks first across three training seeds: mean report perplexity is 1.68639 versus 1.69154 for a matched static-random nonlocal graph, 1.69258 for dense continued training, and 1.81505 for equal-budget local sliding. At seed 1234, the packed-report paired interval places Dense-CT 0.3169-0.4055% above BF1 and static-random graph 17 0.2441-0.3642% above BF1. These results establish BF1 as a reproducible sparse operator and selective retrofit primitive with real long-context systems value. This paper evaluates numerical correctness, selected-interaction scaling, kernel performance, partial-model inference, and matched next-token language modeling.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • arXiv:2608.20427v1 Announce Type: new Abstract: Dense causal attention remains expensive at long context even when implemented with highly optimized exact kernels. We study BF1, a…
站內正文

待翻譯:Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself as an inference cloud after licensing its LPU technology to NVIDIA. This comparison checks each provider's live rate card, Q2 2026 financials, active and contracted gigawatts, anchor contracts, and SemiAnalysis ClusterMAX tier. Nebius posts the lowest H100 rate and the only published B300 price, Lambda has the cheapest B200, Crusoe is the only one with AMD on its card, and CoreWeave commands a 10–15% premium as the sole Platinum-rated provider. Figures verified August 21, 2026. The post Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power appeared first on MarkTechPost.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself…
站內正文

待翻譯:Starcloud raises $250M to build AI data centers in orbit

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Artificial intelligence hardware startup Starcloud Inc. today announced that it has raised $250 million in funding at a $2.3 billion valuation. Investment firm Manhattan West led the deal. It was joined by more than a dozen other backers including Nvidia Corp. and Cisco Investments. The cash infusion extends a Series A round that Starcloud announced in […] The post Starcloud raises $250M to build AI data centers in orbit appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Artificial intelligence hardware startup Starcloud Inc. today announced that it has raised $250 million in funding at a $2.3 billion valuation. Investment firm Manhattan West led…
站內正文

待翻譯:Nvidia just showed that the harness, not the AI model, is now the real hero

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Nvidia published some interesting new research on Friday suggesting it’s the harness, more than the underlying model, that is far more important when asking an AI to do long-horizon tasks. A harness is the software wrap…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Nvidia published some interesting new research on Friday suggesting it’s the harness, more than the underlying model, that is far more important when asking an AI to do long-horiz…
站內正文

待翻譯:Where Security Fits in an AI Agent Stack

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:As AI agents become more capable and operate over longer horizons, building security and trust into the applications they power becomes increasingly important. Drawing on work with NVIDIA OpenShell, agent developers, op…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • As AI agents become more capable and operate over longer horizons, building security and trust into the applications they power becomes increasingly important. Drawing on work wit…
站內正文

待翻譯:Nvidia's Switchyard router reshuffles AI models mid-task to cut task costs

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks to cheaper models and that becomes its…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks…
站內正文

待翻譯:From Retrieved Context to Runtime Control: Adaptive Compression for Edge-based RAG

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2608.19535v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language-model responses by grounding generation in external passages, which comes with overhead: retrieved context lengthens the prompt, increasing prefill work, KV-cache footprint, memory traffic, latency, and energy. Context compression offers a natural remedy by pruning retrieved text before generation. However, state-of-the-art context-compression methods are typically used with a fixed compression budget, or with the rate selected offline and then applied at inference time. This static view ignores both workload variation and the live state of the edge device. On an edge SoC, compression is not free: the compressor itself runs on the same SoC and consumes latency and energy that can offset any generation savings. This paper proposes a vision for telemetry-informed adaptive compression in edge RAG, grounded in experimental evidence. We characterize the compression tradeoff on the NVIDIA Jetson AGX Thor using Llama and Qwen generators, Natural Questions and HotpotQA datasets, and LLMLingua-2 compression. Our measurements show that generation dominates the RAG budget for larger models, reaching roughly 90% of per-query latency and 91% of GPU energy for 7B-8B generators. Exploring the impact of the compression rate reveals an adaptive operating region: mild compression can miss energy opportunities, and overly aggressive compression can hurt inference quality. Intermediate compression can reduce GPU energy by up to 53.2%, and SoC energy by up to 48.2%, with negligible quality loss. We argue for runtime policies that dynamically manage compression, guided by workload features and edge telemetry.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • arXiv:2608.19535v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language-model responses by grounding generation in external passages, which comes wi…
站內正文

待翻譯:Nvidia’s SONIC Teaches Humanoids to Move

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Nvidia’s model uses real-time human demonstrations and training data to give operators a one-stop shop for humanoid motion.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Nvidia’s model uses real-time human demonstrations and training data to give operators a one-stop shop for humanoid motion.
站內正文

待翻譯:Bring the Fire: Play Games on GeForce NOW With New Firefox Browser Support

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:It’s a new way into the cloud. GeForce NOW welcomes Firefox support to the cloud, opening up another way to jump into high-performance PC gaming straight from the browser, starting today. Whether on a school laptop or everyday PC, it’s now even easier to play supported PC games without downloading a dedicated app. Plus, discover […]

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • It’s a new way into the cloud. GeForce NOW welcomes Firefox support to the cloud, opening up another way to jump into high-performance PC gaming straight from the browser, startin…
站內正文

待翻譯:Blowing Off Steam: How Power-Flexible AI Factories Can Stabilize the Global Energy Grid

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Editor’s note: This blog, originally published in March 2026, has been updated. At the half-time whistle of the UEFA EURO 2020 round of 16 football match between England and Germany, millions of viewers stepped away from their screens in the U.K. to do the same thing at the same time — turn on their kettles. […]

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Editor’s note: This blog, originally published in March 2026, has been updated. At the half-time whistle of the UEFA EURO 2020 round of 16 football match between England and Germa…
站內正文

待翻譯:Sanja Fidler’s world model startup Veeda AI raises $90M in seed funding

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Veeda AI, a startup led by a team of former Nvidia Corp. researcher and renowned computer scientist Sanja Fidler, has taken its bow on the main stage after raising $90 million in a seed funding round today. The round, which was first reported by The Logic, was co-led by Khosla Ventures and Radical Ventures, is […] The post Sanja Fidler’s world model startup Veeda AI raises $90M in seed funding appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Veeda AI, a startup led by a team of former Nvidia Corp. researcher and renowned computer scientist Sanja Fidler, has taken its bow on the main stage after raising $90 million in…
站內正文

待翻譯:Nvidia’s new financial strategy does not compute

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:“Compute is an asset class! Compute is an asset class!” I continue to insist as I slowly shrink down and turn into a corncob | Image: Cath Virginia / The Verge, Getty Images April - 1805 Napoleon is master of Europe Only the British fleet stands before him Compute is now an asset class I see it is once again time to talk financial innovation. Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR are all working with Nvidia to put together $500 billion in financing to turn compute into an asset class. "This is really the first time that technology chips have become an investable asset class," Nvidia CEO Jensen Huang said to CNBC. "These are revenue-generating assets now. They're productive, they're long-lived, they're fungible, they're flexible." "This is the very beginning, like what it was whe … Read the full story at The Verge.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • “Compute is an asset class! Compute is an asset class!” I continue to insist as I slowly shrink down and turn into a corncob | Image: Cath Virginia / The Verge, Getty Images April…
站內正文

待翻譯:KernelArc: A Multi-Agent Framework for GPU Kernel Optimization

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:arXiv:2608.17071v1 Announce Type: new Abstract: We present KernelArc, a multi-agent framework for autonomous GPU kernel optimization across heterogeneous workloads. Strategy-specialized agents run in parallel and coordinate through conclusions-only shared memory, a deterministic benchmark guard, and read-only cross-agent state with plateau-triggered drafting. We evaluate \kernelarc{} on NVIDIA H100 and B200 GPUs using category-representative SOL-ExecBench workloads. The resulting implementations span custom BF16 GEMM, static cuBLASLt Expert-API configuration tables, fused mixture-of-experts backward, shape-gated decoder-layer fusion, native NVFP4 grouped-query attention, and paged prefill attention. At the public SOL-ExecBench leaderboard snapshot recorded on July~30, 2026, these submissions ranked first on representative L1, L2, Quantization, and FlashInfer tasks. The trajectories support the paper's central motivation: shared multi-agent search can broaden exploration and reach stronger incumbents within a fixed candidate budget, while the value of individual coordination features depends on the kernel and optimization stage.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • arXiv:2608.17071v1 Announce Type: new Abstract: We present KernelArc, a multi-agent framework for autonomous GPU kernel optimization across heterogeneous workloads. Strategy-speci…
站內正文

待翻譯:NVIDIA Releases TensorRT Model Connect in Public Preview: Hugging Face Checkpoint to Native C++ Inference in Two Commands

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:NVIDIA has released TensorRT Model Connect (TRTMC) in public preview, an Apache-2.0 project that takes a supported Hugging Face or local checkpoint to end-to-end TensorRT inference in two commands, with no intermediate ONNX export. The build emits a versioned .bundle artifact that runs through native C++ task APIs, so inference executes without PyTorch in the runtime path. NVIDIA's July 29, 2026 GB300 snapshot covers 105 release profiles across 76 model families. The post NVIDIA Releases TensorRT Model Connect in Public Preview: Hugging Face Checkpoint to Native C++ Inference in Two Commands appeared first on MarkTechPost.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • NVIDIA has released TensorRT Model Connect (TRTMC) in public preview, an Apache-2.0 project that takes a supported Hugging Face or local checkpoint to end-to-end TensorRT inferenc…
站內正文

待翻譯:Nvidia to Back OpenAI Data Center With $105B Investment

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The financing and infrastructure role puts the AI vendor at the center of another multibillion-dollar data center project.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The financing and infrastructure role puts the AI vendor at the center of another multibillion-dollar data center project.
站內正文

待翻譯:ByteDance Seed and Tsinghua AIR Introduces CUDA Agent: A Large-Scale Agentic RL System for CUDA Kernel Generation

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:ByteDance Seed and Tsinghua AIR have released CUDA Agent, an agentic reinforcement learning system that trains a large language model to write GPU kernels that beat a compiler. The gap it targets is narrow but stubborn: frontier models already produce correct CUDA, they just produce slow CUDA. On KernelBench, the base model Seed1.6 passes 74.0% […] The post ByteDance Seed and Tsinghua AIR Introduces CUDA Agent: A Large-Scale Agentic RL System for CUDA Kernel Generation appeared first on MarkTechPost.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • ByteDance Seed and Tsinghua AIR have released CUDA Agent, an agentic reinforcement learning system that trains a large language model to write GPU kernels that beat a compiler. Th…
站內正文

待翻譯:How NVIDIA scales expertise with ChatGPT Work

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:NVIDIA teams use ChatGPT Work to reduce manual tasks, connect fast-moving signals, and scale successful workflows globally.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • NVIDIA teams use ChatGPT Work to reduce manual tasks, connect fast-moving signals, and scale successful workflows globally.
站內正文

待翻譯:AI cloud operator Groq raises $350M more in funding

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Artificial intelligence startup Groq Inc. today announced that it has raised $350 million in funding. The Series A round was led by returning backer Disruptive. Grok stated that Nvidia Corp. plans to join the round later down the line, but didn’t specify how much the chip giant will invest. The cash infusion comes less than […] The post AI cloud operator Groq raises $350M more in funding appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Artificial intelligence startup Groq Inc. today announced that it has raised $350 million in funding. The Series A round was led by returning backer Disruptive. Grok stated that N…
站內正文

待翻譯:NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture-of-Experts model (3B active), which delivers up to 4x higher throughput and up to 30% faster task completion for always-on agents.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture…
站內正文

待翻譯:On theCUBE Pod: AI bubble debate heats up and neocloud earnings challenge doubters

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The debate over whether we are in an artificial intelligence bubble took a new turn this week. Despite the ballooning AI spending, Dave Vellante (pictured, right), chief analyst for theCUBE Research, contends that any bursting point may be far off. Now that Nvidia Corp. Chief Executive Jensen Huang, has committed $500 billion to establish independent […] The post On theCUBE Pod: AI bubble debate heats up and neocloud earnings challenge doubters appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The debate over whether we are in an artificial intelligence bubble took a new turn this week. Despite the ballooning AI spending, Dave Vellante (pictured, right), chief analyst f…
站內正文

待翻譯:AI-enriched Linux 7.2 delivers cache-aware scheduling - here's everything new

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The latest stable kernel also brings filesystem and I/O improvements, and substantial new support across AMD, Intel, Apple, Nvidia, USB4, and laptop hardware.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The latest stable kernel also brings filesystem and I/O improvements, and substantial new support across AMD, Intel, Apple, Nvidia, USB4, and laptop hardware.
站內正文

待翻譯:Teaching Everyone to Fish for Tokens

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Nvidia wants you building your own model, not buying from Anthropic/OpenAI.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Nvidia wants you building your own model, not buying from Anthropic/OpenAI.
站內正文

待翻譯:LG to Release Nvidia-Powered Humanoid in 2027

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The robot is part of the companies’ expanded partnership to shift physical AI from concepts to real-world deployments.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • The robot is part of the companies’ expanded partnership to shift physical AI from concepts to real-world deployments.
站內正文

待翻譯:Securing the Infrastructure of Intelligence

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:AI factories are the defining infrastructure of the AI era—where compute transforms energy and data into intelligence that powers every business, industry and country. In the AI economy, compute is revenue. AI factories require a full stack of critical resources: advanced chips, packaging, memory, and networking – as well as land, power and shell. Just […]

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • AI factories are the defining infrastructure of the AI era—where compute transforms energy and data into intelligence that powers every business, industry and country. In the AI e…
站內正文

待翻譯:Nvidia AI Financing Is the $500B Risk Investors Aren't Watching

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Nvidia AI Financing Is The $500 Billion Risk Investors Aren’t Watching ByJim Osman, Senior Contributor. Forbes contributors publish independent expert analyses and insights. Jim Osman is a finance expert with over 30 ye…

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Nvidia AI Financing Is The $500 Billion Risk Investors Aren’t Watching ByJim Osman, Senior Contributor. Forbes contributors publish independent expert analyses and insights. Jim O…
站內正文

待翻譯:Did Nvidia’s Jensen Huang just make the AI buildout too big to fail?

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Nvidia Corp. is no longer just selling technology. It is helping create a financial asset class around artificial intelligence compute. In our last Breaking Analysis, we argued that AI can be technologically transformative and still produce a capital bubble. Our thesis was simply that the bubble pops if deployable supply grows faster than monetizable demand – […] The post Did Nvidia’s Jensen Huang just make the AI buildout too big to fail? appeared first on SiliconANGLE.

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Nvidia Corp. is no longer just selling technology. It is helping create a financial asset class around artificial intelligence compute. In our last Breaking Analysis, we argued th…
站內正文

River AI獲11億美元融資:個性化AI革命意味着什麼?

成立僅兩個月的AI初創公司River AI完成11億美元早期融資,創下行業紀錄,Nvidia與AMD罕見聯手參投。其API可在十幾分鍾內基於企業內部數據微調開源模型,使計算成本降低2至4倍,標誌着企業AI由集中式壟斷轉向私有化、定製化,併為中小企業帶來新的GEO/SEO窗口。

  • 成立僅兩個月的River AI完成11億美元融資,是AI行業史上最大早期融資之一,獲Nvidia、AMD、General Catalyst、Y Combinator、淡馬錫等支持。
  • River AI的API允許企業在內部數據上快速微調開源模型,僅需十幾分鍾,計算成本降低2至4倍。
站內正文

UGM、Indosat與NVIDIA為印尼開設首座大學AI中心,培養本土AI人才

本週,印尼通信與數字事務部、Indosat Ooredoo Hutchison、NVIDIA與加查馬達大學在日惹共同啓動了UGM Indosat NVIDIA AI技術中心(NVAITC),這是該國首個大學級AI中心。該中心依託印尼AI卓越中心計劃,配備NVIDIA全棧AI平台和Indosat的GPU Merdeka主權GPU即服務平台,聚焦醫療、農業和災害應對三個初始項目,旨在培養本土AI人才,推動印尼從AI使用者轉變為AI創新貢獻者。

  • 該中心是印尼首個大學級AI技術中心,由Komdigi、Indosat、NVIDIA和UGM合作成立。
  • 配備NVIDIA全棧AI平台和Indosat主權GPU即服務,支持研究人員和學生使用企業級算力與AI工具。
站內正文

NVIDIA Nemotron 3.5 Lightning:AI智能體的高效執行引擎

NVIDIA Nemotron 3.5 Lightning 是專為AI智能體執行層設計的開源權重模型,擁有30B總參數和3B激活參數,採用混合Mamba-2+MoE+注意力架構,支持最高1M上下文。它旨在讓昂貴的前沿模型負責規劃,自身承擔高頻工具調用等執行工作,據稱輸出速度可達同類模型的4倍,並通過多個渠道提供免費或低成本訪問。

  • Open-weight,30B總參數/3B激活,混合Mamba-2+MoE+選擇性注意力架構,上下文最高1M token。
  • 面向智能體執行層,用快速模型處理高頻工具調用、驗證、命令等,降低成本和延遲。
站內正文

公司導航

NVIDIA — AI 公司追蹤 | AI News Hub