<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Wan2.2 on Text Matrix</title><link>https://txtmix.com/tags/wan2.2/</link><description>Recent content in Wan2.2 on Text Matrix</description><generator>Hugo</generator><language>zh-cn</language><lastBuildDate>Tue, 21 Jul 2026 20:06:14 +0800</lastBuildDate><atom:link href="https://txtmix.com/tags/wan2.2/index.xml" rel="self" type="application/rss+xml"/><item><title>Bernini 拆解：字节跳动把 MLLM 语义规划器和 DiT 渲染器拆开，到底在解决什么问题</title><link>https://txtmix.com/posts/tech/bytedance-bernini-mllm-dit-video-generation-architecture/</link><pubDate>Fri, 05 Jun 2026 09:30:00 +0800</pubDate><guid>https://txtmix.com/posts/tech/bytedance-bernini-mllm-dit-video-generation-architecture/</guid><description>Bernini 不是又一个 DiT 视频模型。它把「MLLM 语义规划 + Wan2.2 双专家 DiT 渲染 + Open-VeOmni 序列并行」三段式架构开源，并且把 6 类视频任务（t2i/i2i/t2v/v2v/mv2v/rv2v/r2v）和 7 种 guidance mode 显式化。本文从 Bernini 仓库的 configs、pipeline.py、parallel/ops.py 三个核心文件出发，拆出这套架构的设计取舍与适用边界。</description></item></channel></rss>