跳到主要内容
AI News HubLIVE
更多
来源内容 · 翻译待补全1 分钟阅读

待翻译:Modal Clusters are generally available

文章摘要

AI 服务暂时不可用,以下为来源摘要,待恢复后补全翻译:Multi-node GPU clusters with RDMA, gang scheduled from Modal's shared capacity pool and billed by the second, behind a single decorator.

来源Modal Blog作者: Peyton Walters
待翻译:Modal Clusters are generally available
报告错误

纠错通道尚未开通,可先复制下方文章信息留存。

查看更正说明
直接读正文

AI 服务暂时不可用,以下为来源正文,待恢复后补全翻译。

Organizations need to own their intelligence to be successful. Training and serving that intelligence at scale, however, requires petaFLOP/s of compute and terabit/s of networking, spread across many nodes. But owning intelligence doesn’t need to mean owning that hardware. For the past 1.5 years, we’ve been battle-testing a new primitive: Modal Clusters. Today, we’re excited to announce that they are generally available through a single decorator, @modal.clustered: @app.function(gpu="B300:8") @modal.clustered(size=4, rdma=True) def train_model(): cluster = modal.Cluster.from_context() container_ips = cluster.private_ips() container_rank = cluster.container_rank() world_size = len(container_ips) main_addr = container_ips[0] print(f"{container_rank=} {world_size=} {main_addr=}") ...

展开要点与分析

文章情报

工程师进阶

要点

  • AI 服务暂时不可用,系统已先保留来源内容与降级元数据。
  • Multi-node GPU clusters with RDMA, gang scheduled from Modal's shared capacity pool and billed by the second, behind a single decorator.

要点与分析由自动化流程生成,可能有误,请结合原始来源核实。