AI News HubLIVE

来源分布

  • SiliconANGLE AI12
  • Hacker News AI10
  • NVIDIA Blog7
  • MarkTechPost5
  • AI Business3
  • arXiv AI2
  • arXiv Machine Learning1
  • AWS Machine Learning Blog1

主题分布

  • 芯片49
  • Agent26
  • 研究10
  • 模型9
  • 创业融资8
  • 机器人3
  • 工具1

日期线

  • 2026-08-259
  • 2026-08-187
  • 2026-08-266
  • 2026-08-175
  • 2026-08-245
  • 2026-08-194
  • 2026-08-204
  • 2026-08-224

最新动态

待翻译:Nvidia Jetson Orin-guided Russian AI drone killed three civilians in Ukraine

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:5 Join the conversation Follow us Add us as a preferred source on Google A Russian Molniya drone carrying an Nvidia Jetson Orin module crashed and killed three civilians at a gas station in Zaporizhzhia last month after…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • 5 Join the conversation Follow us Add us as a preferred source on Google A Russian Molniya drone carrying an Nvidia Jetson Orin module crashed and killed three civilians at a gas…
站内正文

待翻译:Perplexity's Portable Computer tackling local AI services market

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Now Perplexity is trying to get into the local AI action Amid talk of an Nvidia deal, the AI search biz is looking beyond the cloud Thomas Claburn Thomas Claburn AI AND SOFTWARE REPORTER Published wed 26 Aug 2026 // 00:…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Now Perplexity is trying to get into the local AI action Amid talk of an Nvidia deal, the AI search biz is looking beyond the cloud Thomas Claburn Thomas Claburn AI AND SOFTWARE R…
站内正文

待翻译:Perplexity AI launches Portable Computer on-device AI agent

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Perplexity AI Inc. today introduced Portable Computer, an artificial intelligence agent designed to run on desktops equipped with Nvidia Corp. silicon. The launch follows a report that Nvidia is weighing an investment in the startup that could value it at over $30 billion. Furthermore, Nvidia has reportedly floated the idea of licensing Perplexity’s technology and […] The post Perplexity AI launches Portable Computer on-device AI agent appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Perplexity AI Inc. today introduced Portable Computer, an artificial intelligence agent designed to run on desktops equipped with Nvidia Corp. silicon. The launch follows a report…
站内正文

待翻译:Perplexity Ships Portable Computer on NVIDIA DGX Spark: Local Harness, OS-Enforced Sandbox, and Zero Per-Token Cost for Local Steps

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Perplexity releases Portable Computer, packaging local models, harness, sandbox, and connectors into one system running on NVIDIA DGX Spark. The post Perplexity Ships Portable Computer on NVIDIA DGX Spark: Local Harness, OS-Enforced Sandbox, and Zero Per-Token Cost for Local Steps appeared first on MarkTechPost.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Perplexity releases Portable Computer, packaging local models, harness, sandbox, and connectors into one system running on NVIDIA DGX Spark. The post Perplexity Ships Portable Com…
站内正文

待翻译:Taiwan charges nine people for smuggling ‘high-end’ AI servers to China

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Among those charged are two Super Micro employees and one from Nvidia, marking another flashpoint in US-China AI rivalry Taiwanese prosecutors charged nine people Monday, including one from Nvidia and two from Super Micro, for illegally exporting “high-end AI servers” to mainland China, adding another wave of turbulence in the AI ​​rivalry between China and the United States. Prosecutors said the servers involved were graphics processing units known as “B300,” which have been banned from sale to China. Continue reading...

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Among those charged are two Super Micro employees and one from Nvidia, marking another flashpoint in US-China AI rivalry Taiwanese prosecutors charged nine people Monday, includin…
站内正文

待翻译:Nvidia's Groq 3 LPX accelerator enters full production at 3,500 tokens/SEC

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production as it strives to maintain its dominance in the world of AI compute. The new chip, announ…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production as it strives to maintain its dominance in the w…
站内正文

待翻译:Perplexity’s Computer agent can now run locally — if you can afford it

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Perplexity, in partnership with Nvidia, has taken Computer, its agentic AI assistant, and brought it to the desktop in the The post Perplexity’s Computer agent can now run locally — if you can afford it appeared first on The New Stack.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Perplexity, in partnership with Nvidia, has taken Computer, its agentic AI assistant, and brought it to the desktop in the The post Perplexity’s Computer agent can now run locally…
站内正文

待翻译:Leading Publishers Bring Blockbuster PC Games and Technology to NVIDIA RTX Spark

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:NVIDIA is bringing the next wave of RTX gaming to the Gamescom conference running this week in Cologne, Germany, with support for new games, anti-cheat technologies and increased visual quality. Electronic Arts, Embark and Ubisoft are among the latest game publishers and developers bringing their blockbuster titles to NVIDIA RTX Spark ahead of its launch […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • NVIDIA is bringing the next wave of RTX gaming to the Gamescom conference running this week in Cologne, Germany, with support for new games, anti-cheat technologies and increased…
站内正文

待翻译:Nvidia doubles compute for entry-level edge robotics with Jetson Orin Nano 2

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Nvidia Corp. today announced the release of Jetson Orin Nano 2, a robotics computer “brain” for running artificial intelligence and frontier-level models at the edge. In the past months, foundational AI models have grown smaller and more efficient, adding numerous capabilities alongside language understanding, computer vision and audio processing. As more AI models compress in […] The post Nvidia doubles compute for entry-level edge robotics with Jetson Orin Nano 2 appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Nvidia Corp. today announced the release of Jetson Orin Nano 2, a robotics computer “brain” for running artificial intelligence and frontier-level models at the edge. In the past…
站内正文

待翻译:AI factories enter the execution era as Cisco and NVIDIA push rack-scale systems into production

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The artificial intelligence infrastructure market is crossing an important threshold. The conversation is shifting from acquiring graphics processing units to building complete AI factories that can generate tokens reliably, efficiently and at scale. This is the next bottleneck. GPUs may be the engine, but an AI factory is a system. Compute, networking, storage, cooling, software […] The post AI factories enter the execution era as Cisco and NVIDIA push rack-scale systems into production appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The artificial intelligence infrastructure market is crossing an important threshold. The conversation is shifting from acquiring graphics processing units to building complete AI…
站内正文

待翻译:A Go dependency wrote AGENTS.md mid-build and got Codex to hide the change

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:In a proof of concept published in mid-2026, NVIDIA's AI Red Team planted a malicious dependency in a Go project, let it execute during a normal build, and watched it write a new AGENTS.md file into the project director…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • In a proof of concept published in mid-2026, NVIDIA's AI Red Team planted a malicious dependency in a Go project, let it execute during a normal build, and watched it write a new…
站内正文

待翻译:Nvidia and Cisco push the enterprise AI factory into the rack-scale era

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The Cisco Secure AI Factory with Nvidia has been extended into the rack-scale era, offering enterprises greater full-stack operational capabilities as a result. Designed to give customers a framework for deploying artificial intelligence across their entire infrastructure, Secure AI Factory with Nvidia from Cisco Systems Inc. now integrates rack-to-fabric liquid cooling supporting systems beyond 200 […] The post Nvidia and Cisco push the enterprise AI factory into the rack-scale era appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The Cisco Secure AI Factory with Nvidia has been extended into the rack-scale era, offering enterprises greater full-stack operational capabilities as a result. Designed to give c…
站内正文

待翻译:Cisco and Nvidia take AI factories from rack to runtime

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:AI factories are moving from ambitious plans toward production, but the path from graphics processing unit acquisition to usable systems remains a race against time. Neoclouds already have customers waiting for capacity, enterprises are looking to bring inference workloads closer to home and sovereign AI programs are being built now. Those distinct buyer motions are […] The post Cisco and Nvidia take AI factories from rack to runtime appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • AI factories are moving from ambitious plans toward production, but the path from graphics processing unit acquisition to usable systems remains a race against time. Neoclouds alr…
站内正文

待翻译:Robotics AI startup Generalist reportedly raises $200M

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Generalist AI Inc., a startup that develops artificial intelligence software for robots, has reportedly raised $200 million in funding. Axios today cited a source as saying that 8VC led the investment. It was reportedly joined by a number of unnamed existing investors. Generalist’s previous $400 million round in June included the participation of Nvidia Corp., […] The post Robotics AI startup Generalist reportedly raises $200M appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Generalist AI Inc., a startup that develops artificial intelligence software for robots, has reportedly raised $200 million in funding. Axios today cited a source as saying that 8…
站内正文

待翻译:Nvidia reportedly eyes another investment in Perplexity AI at a $30B valuation

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Nvidia Corp. is reportedly considering making another investment in the artificial intelligence search startup Perplexity AI Inc. A report by The Information says the chipmaker is holding talks with Perplexity over an investment that could push the startup’s valuation to more than $30 billion. That would represent a jump of more than 50% from the […] The post Nvidia reportedly eyes another investment in Perplexity AI at a $30B valuation appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Nvidia Corp. is reportedly considering making another investment in the artificial intelligence search startup Perplexity AI Inc. A report by The Information says the chipmaker is…
站内正文

待翻译:How XPUs Meet a World-Class AI Factory

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:To generate intelligence at scale, AI factories run continuously, and their economics are defined by delivered output: tokens per second, tokens per watt, cost per token, utilization and uptime. That requires AI infrastructure designed and built as a full factory, not a collection of individual accelerators. Hyperscalers and AI-native companies building custom XPUs must consider […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • To generate intelligence at scale, AI factories run continuously, and their economics are defined by delivered output: tokens per second, tokens per watt, cost per token, utilizat…
站内正文

待翻译:With Groq 3 LPX in Full Production, NVIDIA Extends Vera Rubin Inference for Agents

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The next era of AI inference won’t be defined by a single breakthrough chip, network or system. It’ll be defined by how every layer of the AI factory works together. That’s why NVIDIA is extending Vera Rubin NVL72 with fast token generation for agentic systems. Announced today, the NVIDIA Vera Rubin rack-scale system NVIDIA Groq […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The next era of AI inference won’t be defined by a single breakthrough chip, network or system. It’ll be defined by how every layer of the AI factory works together. That’s why NV…
站内正文

待翻译:Up to 30x More Work Per Watt: NVIDIA Vera Rubin NVL72 Sets a New Efficiency Standard for AI Agents

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:According to OpenRouter data, agentic AI workloads consume 15x more tokens than a simple chat request. Why? Consider what happens when an AI agent researches a company for an investment decision. The agent queries financial databases, searches news and filings, invokes a sub-agent to run peer comparisons and model valuations, then synthesizes everything into a […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • According to OpenRouter data, agentic AI workloads consume 15x more tokens than a simple chat request. Why? Consider what happens when an AI agent researches a company for an inve…
站内正文

待翻译:Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself as an inference cloud after licensing its LPU technology to NVIDIA. This comparison checks each provider's live rate card, Q2 2026 financials, active and contracted gigawatts, anchor contracts, and SemiAnalysis ClusterMAX tier. Nebius posts the lowest H100 rate and the only published B300 price, Lambda has the cheapest B200, Crusoe is the only one with AMD on its card, and CoreWeave commands a 10–15% premium as the sole Platinum-rated provider. Figures verified August 21, 2026. The post Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power appeared first on MarkTechPost.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself…
站内正文

待翻译:BF1: A Causal Dyadic Sparse-Attention Retrofit for Efficient Long-Context Transformers

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:arXiv:2608.20427v1 Announce Type: new Abstract: Dense causal attention remains expensive at long context even when implemented with highly optimized exact kernels. We study BF1, a deterministic block-aligned dyadic sparse-attention route that combines a small exact local neighborhood, a global first block, and logarithmically spaced historical blocks. The route is related to prior log-sparse and dilated attention patterns; our contribution is a correctness-gated pretrained-model retrofit, a matched topology-control study, and a systems characterization that connects per-layer sparsity to whole-model latency. For fixed block width, every converted layer uses O(n log n) selected token interactions and has O(log n) graph communication depth. On an NVIDIA RTX PRO 6000 Blackwell GPU, an optimized BF16 implementation crosses dense attention between 2K and 4K tokens and reaches a 10.91x per-layer prefill speedup at 32K. Retrofitting eight of 28 Qwen3-0.6B attention layers lowers warm whole-model time to first token by 7.7%, 11.3%, and 15.3% at 8K, 16K, and 32K, respectively, while the remaining dense layers keep the complete model asymptotically quadratic. Under a matched 1,000-step, 16.384M-token adaptation protocol, BF1 ranks first across three training seeds: mean report perplexity is 1.68639 versus 1.69154 for a matched static-random nonlocal graph, 1.69258 for dense continued training, and 1.81505 for equal-budget local sliding. At seed 1234, the packed-report paired interval places Dense-CT 0.3169-0.4055% above BF1 and static-random graph 17 0.2441-0.3642% above BF1. These results establish BF1 as a reproducible sparse operator and selective retrofit primitive with real long-context systems value. This paper evaluates numerical correctness, selected-interaction scaling, kernel performance, partial-model inference, and matched next-token language modeling.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • arXiv:2608.20427v1 Announce Type: new Abstract: Dense causal attention remains expensive at long context even when implemented with highly optimized exact kernels. We study BF1, a…
站内正文

待翻译:Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself as an inference cloud after licensing its LPU technology to NVIDIA. This comparison checks each provider's live rate card, Q2 2026 financials, active and contracted gigawatts, anchor contracts, and SemiAnalysis ClusterMAX tier. Nebius posts the lowest H100 rate and the only published B300 price, Lambda has the cheapest B200, Crusoe is the only one with AMD on its card, and CoreWeave commands a 10–15% premium as the sole Platinum-rated provider. Figures verified August 21, 2026. The post Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power appeared first on MarkTechPost.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The five largest GPU neoclouds now run on very different models. CoreWeave and Nebius report to the SEC; Lambda and Crusoe are private and heading toward IPOs; Groq rebuilt itself…
站内正文

待翻译:Starcloud raises $250M to build AI data centers in orbit

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Artificial intelligence hardware startup Starcloud Inc. today announced that it has raised $250 million in funding at a $2.3 billion valuation. Investment firm Manhattan West led the deal. It was joined by more than a dozen other backers including Nvidia Corp. and Cisco Investments. The cash infusion extends a Series A round that Starcloud announced in […] The post Starcloud raises $250M to build AI data centers in orbit appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Artificial intelligence hardware startup Starcloud Inc. today announced that it has raised $250 million in funding at a $2.3 billion valuation. Investment firm Manhattan West led…
站内正文

待翻译:Nvidia just showed that the harness, not the AI model, is now the real hero

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Nvidia published some interesting new research on Friday suggesting it’s the harness, more than the underlying model, that is far more important when asking an AI to do long-horizon tasks. A harness is the software wrap…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Nvidia published some interesting new research on Friday suggesting it’s the harness, more than the underlying model, that is far more important when asking an AI to do long-horiz…
站内正文

待翻译:Where Security Fits in an AI Agent Stack

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:As AI agents become more capable and operate over longer horizons, building security and trust into the applications they power becomes increasingly important. Drawing on work with NVIDIA OpenShell, agent developers, op…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • As AI agents become more capable and operate over longer horizons, building security and trust into the applications they power becomes increasingly important. Drawing on work wit…
站内正文

待翻译:Nvidia's Switchyard router reshuffles AI models mid-task to cut task costs

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks to cheaper models and that becomes its…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Enterprises running always-on AI agents keep hitting the same tradeoff. Send every task to a frontier model and the bill climbs fast. Build custom routing logic to send easy tasks…
站内正文

待翻译:From Retrieved Context to Runtime Control: Adaptive Compression for Edge-based RAG

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:arXiv:2608.19535v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language-model responses by grounding generation in external passages, which comes with overhead: retrieved context lengthens the prompt, increasing prefill work, KV-cache footprint, memory traffic, latency, and energy. Context compression offers a natural remedy by pruning retrieved text before generation. However, state-of-the-art context-compression methods are typically used with a fixed compression budget, or with the rate selected offline and then applied at inference time. This static view ignores both workload variation and the live state of the edge device. On an edge SoC, compression is not free: the compressor itself runs on the same SoC and consumes latency and energy that can offset any generation savings. This paper proposes a vision for telemetry-informed adaptive compression in edge RAG, grounded in experimental evidence. We characterize the compression tradeoff on the NVIDIA Jetson AGX Thor using Llama and Qwen generators, Natural Questions and HotpotQA datasets, and LLMLingua-2 compression. Our measurements show that generation dominates the RAG budget for larger models, reaching roughly 90% of per-query latency and 91% of GPU energy for 7B-8B generators. Exploring the impact of the compression rate reveals an adaptive operating region: mild compression can miss energy opportunities, and overly aggressive compression can hurt inference quality. Intermediate compression can reduce GPU energy by up to 53.2%, and SoC energy by up to 48.2%, with negligible quality loss. We argue for runtime policies that dynamically manage compression, guided by workload features and edge telemetry.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • arXiv:2608.19535v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language-model responses by grounding generation in external passages, which comes wi…
站内正文

待翻译:Nvidia’s SONIC Teaches Humanoids to Move

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Nvidia’s model uses real-time human demonstrations and training data to give operators a one-stop shop for humanoid motion.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Nvidia’s model uses real-time human demonstrations and training data to give operators a one-stop shop for humanoid motion.
站内正文

待翻译:Bring the Fire: Play Games on GeForce NOW With New Firefox Browser Support

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:It’s a new way into the cloud. GeForce NOW welcomes Firefox support to the cloud, opening up another way to jump into high-performance PC gaming straight from the browser, starting today. Whether on a school laptop or everyday PC, it’s now even easier to play supported PC games without downloading a dedicated app. Plus, discover […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • It’s a new way into the cloud. GeForce NOW welcomes Firefox support to the cloud, opening up another way to jump into high-performance PC gaming straight from the browser, startin…
站内正文

待翻译:Blowing Off Steam: How Power-Flexible AI Factories Can Stabilize the Global Energy Grid

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Editor’s note: This blog, originally published in March 2026, has been updated. At the half-time whistle of the UEFA EURO 2020 round of 16 football match between England and Germany, millions of viewers stepped away from their screens in the U.K. to do the same thing at the same time — turn on their kettles. […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Editor’s note: This blog, originally published in March 2026, has been updated. At the half-time whistle of the UEFA EURO 2020 round of 16 football match between England and Germa…
站内正文

待翻译:Sanja Fidler’s world model startup Veeda AI raises $90M in seed funding

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Veeda AI, a startup led by a team of former Nvidia Corp. researcher and renowned computer scientist Sanja Fidler, has taken its bow on the main stage after raising $90 million in a seed funding round today. The round, which was first reported by The Logic, was co-led by Khosla Ventures and Radical Ventures, is […] The post Sanja Fidler’s world model startup Veeda AI raises $90M in seed funding appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Veeda AI, a startup led by a team of former Nvidia Corp. researcher and renowned computer scientist Sanja Fidler, has taken its bow on the main stage after raising $90 million in…
站内正文

待翻译:Nvidia’s new financial strategy does not compute

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:“Compute is an asset class! Compute is an asset class!” I continue to insist as I slowly shrink down and turn into a corncob | Image: Cath Virginia / The Verge, Getty Images April - 1805 Napoleon is master of Europe Only the British fleet stands before him Compute is now an asset class I see it is once again time to talk financial innovation. Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR are all working with Nvidia to put together $500 billion in financing to turn compute into an asset class. "This is really the first time that technology chips have become an investable asset class," Nvidia CEO Jensen Huang said to CNBC. "These are revenue-generating assets now. They're productive, they're long-lived, they're fungible, they're flexible." "This is the very beginning, like what it was whe … Read the full story at The Verge.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • “Compute is an asset class! Compute is an asset class!” I continue to insist as I slowly shrink down and turn into a corncob | Image: Cath Virginia / The Verge, Getty Images April…
站内正文

待翻译:KernelArc: A Multi-Agent Framework for GPU Kernel Optimization

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:arXiv:2608.17071v1 Announce Type: new Abstract: We present KernelArc, a multi-agent framework for autonomous GPU kernel optimization across heterogeneous workloads. Strategy-specialized agents run in parallel and coordinate through conclusions-only shared memory, a deterministic benchmark guard, and read-only cross-agent state with plateau-triggered drafting. We evaluate \kernelarc{} on NVIDIA H100 and B200 GPUs using category-representative SOL-ExecBench workloads. The resulting implementations span custom BF16 GEMM, static cuBLASLt Expert-API configuration tables, fused mixture-of-experts backward, shape-gated decoder-layer fusion, native NVFP4 grouped-query attention, and paged prefill attention. At the public SOL-ExecBench leaderboard snapshot recorded on July~30, 2026, these submissions ranked first on representative L1, L2, Quantization, and FlashInfer tasks. The trajectories support the paper's central motivation: shared multi-agent search can broaden exploration and reach stronger incumbents within a fixed candidate budget, while the value of individual coordination features depends on the kernel and optimization stage.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • arXiv:2608.17071v1 Announce Type: new Abstract: We present KernelArc, a multi-agent framework for autonomous GPU kernel optimization across heterogeneous workloads. Strategy-speci…
站内正文

待翻译:NVIDIA Releases TensorRT Model Connect in Public Preview: Hugging Face Checkpoint to Native C++ Inference in Two Commands

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:NVIDIA has released TensorRT Model Connect (TRTMC) in public preview, an Apache-2.0 project that takes a supported Hugging Face or local checkpoint to end-to-end TensorRT inference in two commands, with no intermediate ONNX export. The build emits a versioned .bundle artifact that runs through native C++ task APIs, so inference executes without PyTorch in the runtime path. NVIDIA's July 29, 2026 GB300 snapshot covers 105 release profiles across 76 model families. The post NVIDIA Releases TensorRT Model Connect in Public Preview: Hugging Face Checkpoint to Native C++ Inference in Two Commands appeared first on MarkTechPost.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • NVIDIA has released TensorRT Model Connect (TRTMC) in public preview, an Apache-2.0 project that takes a supported Hugging Face or local checkpoint to end-to-end TensorRT inferenc…
站内正文

待翻译:Nvidia to Back OpenAI Data Center With $105B Investment

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The financing and infrastructure role puts the AI vendor at the center of another multibillion-dollar data center project.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The financing and infrastructure role puts the AI vendor at the center of another multibillion-dollar data center project.
站内正文

待翻译:ByteDance Seed and Tsinghua AIR Introduces CUDA Agent: A Large-Scale Agentic RL System for CUDA Kernel Generation

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:ByteDance Seed and Tsinghua AIR have released CUDA Agent, an agentic reinforcement learning system that trains a large language model to write GPU kernels that beat a compiler. The gap it targets is narrow but stubborn: frontier models already produce correct CUDA, they just produce slow CUDA. On KernelBench, the base model Seed1.6 passes 74.0% […] The post ByteDance Seed and Tsinghua AIR Introduces CUDA Agent: A Large-Scale Agentic RL System for CUDA Kernel Generation appeared first on MarkTechPost.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • ByteDance Seed and Tsinghua AIR have released CUDA Agent, an agentic reinforcement learning system that trains a large language model to write GPU kernels that beat a compiler. Th…
站内正文

待翻译:How NVIDIA scales expertise with ChatGPT Work

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:NVIDIA teams use ChatGPT Work to reduce manual tasks, connect fast-moving signals, and scale successful workflows globally.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • NVIDIA teams use ChatGPT Work to reduce manual tasks, connect fast-moving signals, and scale successful workflows globally.
站内正文

待翻译:AI cloud operator Groq raises $350M more in funding

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Artificial intelligence startup Groq Inc. today announced that it has raised $350 million in funding. The Series A round was led by returning backer Disruptive. Grok stated that Nvidia Corp. plans to join the round later down the line, but didn’t specify how much the chip giant will invest. The cash infusion comes less than […] The post AI cloud operator Groq raises $350M more in funding appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Artificial intelligence startup Groq Inc. today announced that it has raised $350 million in funding. The Series A round was led by returning backer Disruptive. Grok stated that N…
站内正文

待翻译:NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture-of-Experts model (3B active), which delivers up to 4x higher throughput and up to 30% faster task completion for always-on agents.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture…
站内正文

待翻译:On theCUBE Pod: AI bubble debate heats up and neocloud earnings challenge doubters

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The debate over whether we are in an artificial intelligence bubble took a new turn this week. Despite the ballooning AI spending, Dave Vellante (pictured, right), chief analyst for theCUBE Research, contends that any bursting point may be far off. Now that Nvidia Corp. Chief Executive Jensen Huang, has committed $500 billion to establish independent […] The post On theCUBE Pod: AI bubble debate heats up and neocloud earnings challenge doubters appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The debate over whether we are in an artificial intelligence bubble took a new turn this week. Despite the ballooning AI spending, Dave Vellante (pictured, right), chief analyst f…
站内正文

待翻译:AI-enriched Linux 7.2 delivers cache-aware scheduling - here's everything new

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The latest stable kernel also brings filesystem and I/O improvements, and substantial new support across AMD, Intel, Apple, Nvidia, USB4, and laptop hardware.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The latest stable kernel also brings filesystem and I/O improvements, and substantial new support across AMD, Intel, Apple, Nvidia, USB4, and laptop hardware.
站内正文

待翻译:Teaching Everyone to Fish for Tokens

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Nvidia wants you building your own model, not buying from Anthropic/OpenAI.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Nvidia wants you building your own model, not buying from Anthropic/OpenAI.
站内正文

待翻译:LG to Release Nvidia-Powered Humanoid in 2027

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:The robot is part of the companies’ expanded partnership to shift physical AI from concepts to real-world deployments.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • The robot is part of the companies’ expanded partnership to shift physical AI from concepts to real-world deployments.
站内正文

待翻译:Securing the Infrastructure of Intelligence

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:AI factories are the defining infrastructure of the AI era—where compute transforms energy and data into intelligence that powers every business, industry and country. In the AI economy, compute is revenue. AI factories require a full stack of critical resources: advanced chips, packaging, memory, and networking – as well as land, power and shell. Just […]

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • AI factories are the defining infrastructure of the AI era—where compute transforms energy and data into intelligence that powers every business, industry and country. In the AI e…
站内正文

待翻译:Nvidia AI Financing Is the $500B Risk Investors Aren't Watching

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Nvidia AI Financing Is The $500 Billion Risk Investors Aren’t Watching ByJim Osman, Senior Contributor. Forbes contributors publish independent expert analyses and insights. Jim Osman is a finance expert with over 30 ye…

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Nvidia AI Financing Is The $500 Billion Risk Investors Aren’t Watching ByJim Osman, Senior Contributor. Forbes contributors publish independent expert analyses and insights. Jim O…
站内正文

俄导弹被曝使用英伟达AI芯片协助打击乌克兰

乌克兰军事情报局在俄军S-71“单色”巡航导弹残骸中发现英伟达Jetson Orin模块,表明该导弹可能使用人工智能技术。尽管英伟达2022年已退出俄罗斯并受制裁,芯片仍通过其他渠道流入俄方,显示现有出口管制无效,需加强制裁压力。

  • 乌军在俄S-71M巡航导弹残骸中发现英伟达Jetson Orin NX 16GB模块。
  • 该芯片支持导弹自主搜索与攻击目标,可能用于AI制导。
站内正文

CPU复兴时代来临

随着智能体AI的兴起,曾被边缘化的CPU重新成为算力瓶颈。从AWS要求工程师节省CPU资源,到Intel、AMD、Arm、高通和英伟达纷纷加码服务器CPU,行业正面临一轮由智能体工作负载驱动的CPU需求激增。

  • 智能体AI大量依赖CPU执行工具调用、安全护栏和文本分词等任务,导致CPU需求骤增。
  • AWS已要求工程师不惜代价节省CPU周期,服务器CPU等待时间激增。
站内正文

待翻译:Did Nvidia’s Jensen Huang just make the AI buildout too big to fail?

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Nvidia Corp. is no longer just selling technology. It is helping create a financial asset class around artificial intelligence compute. In our last Breaking Analysis, we argued that AI can be technologically transformative and still produce a capital bubble. Our thesis was simply that the bubble pops if deployable supply grows faster than monetizable demand – […] The post Did Nvidia’s Jensen Huang just make the AI buildout too big to fail? appeared first on SiliconANGLE.

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Nvidia Corp. is no longer just selling technology. It is helping create a financial asset class around artificial intelligence compute. In our last Breaking Analysis, we argued th…
站内正文

River AI获11亿美元融资:个性化AI革命意味着什么?

成立仅两个月的AI初创公司River AI完成11亿美元早期融资,创下行业纪录,Nvidia与AMD罕见联手参投。其API可在十几分钟内基于企业内部数据微调开源模型,使计算成本降低2至4倍,标志着企业AI由集中式垄断转向私有化、定制化,并为中小企业带来新的GEO/SEO窗口。

  • 成立仅两个月的River AI完成11亿美元融资,是AI行业史上最大早期融资之一,获Nvidia、AMD、General Catalyst、Y Combinator、淡马锡等支持。
  • River AI的API允许企业在内部数据上快速微调开源模型,仅需十几分钟,计算成本降低2至4倍。
站内正文

公司导航

NVIDIA — AI 公司追踪 | AI News Hub