<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>并行策略 on Yiwen Cai</title><link>https://yiwen-cai.github.io/tags/%E5%B9%B6%E8%A1%8C%E7%AD%96%E7%95%A5/</link><description>Recent content in 并行策略 on Yiwen Cai</description><generator>Hugo -- gohugo.io</generator><language>zh-cn</language><managingEditor>caiyiwen.cs@foxmail.com (Yiwen Cai)</managingEditor><webMaster>caiyiwen.cs@foxmail.com (Yiwen Cai)</webMaster><copyright>© 2026 Yiwen Cai</copyright><lastBuildDate>Fri, 10 Jul 2026 21:12:36 +0800</lastBuildDate><atom:link href="https://yiwen-cai.github.io/tags/%E5%B9%B6%E8%A1%8C%E7%AD%96%E7%95%A5/index.xml" rel="self" type="application/rss+xml"/><item><title>分布式训练并行策略：CS336 Lecture 7 笔记</title><link>https://yiwen-cai.github.io/notes/systems/cs336-distributed-parallelism/</link><pubDate>Mon, 15 Jun 2026 00:00:00 +0000</pubDate><author>caiyiwen.cs@foxmail.com (Yiwen Cai)</author><guid>https://yiwen-cai.github.io/notes/systems/cs336-distributed-parallelism/</guid><description>从单 GPU 扩展到多 GPU/多机并行：集合通信原语（all-reduce = reduce-scatter + all-gather）、NVLink/InfiniBand 互联，以及 DDP、FSDP/ZeRO、Tensor/Pipeline/Sequence Parallelism 的取舍与实践法则。</description></item></channel></rss>