Skip to content

Commit e48ae4d

Browse files
committed
Refine changelog content for accuracy and clarity
- Update v0.3.55 to focus specifically on Wan2.2 S2V features - Simplify model descriptions to reflect actual functionality - Remove redundant marketing language - Clarify technical details for better user understanding - Streamline template and API descriptions - Apply same improvements to both English and Chinese versions
1 parent 3575563 commit e48ae4d

2 files changed

Lines changed: 26 additions & 34 deletions

File tree

changelog/index.mdx

Lines changed: 13 additions & 16 deletions
Original file line numberDiff line numberDiff line change
@@ -6,26 +6,25 @@ icon: "clock-rotate-left"
66

77
<Update label="v0.3.55" description="August 29, 2025">
88

9-
**Sound-to-Video Workflow Enhancements & Model Support Expansion**
9+
**Wan2.2 S2V Workflow Enhancements & Model Support Expansion**
1010

11-
This release focuses on improving audio-visual workflow capabilities and expanding model support, particularly for sound-to-video generation and the latest Wan2.2 5B models:
11+
This release focuses on Wan2.2 S2V related video workflow capabilities and model support expansion:
1212

13-
**Sound-to-Video Workflow Control**
14-
- **WanSoundImageToVideoExtend Node**: New manual video extension node for sound-to-video workflows, giving creators precise control over generated video length and timing. This enables fine-tuned control over how audio content translates to video sequences.
13+
**Wan2.2 S2V Workflow Control**
14+
- **WanSoundImageToVideoExtend Node**: New manual video extension node for audio-driven video workflows, giving creators precise control over generated video length and timing. This enables fine-tuned control over how audio content translates to video sequences.
1515
- **Audio-Video Synchronization**: Fixed critical issue where extending video past audio length caused workflow failures, ensuring reliable sound-to-video generation regardless of audio duration.
1616
- **Automatic Audio Trimming**: Video saves now automatically trim audio to match video length, eliminating audio-video sync issues in final output files.
1717

1818
**Advanced Latent Processing**
19-
- **LatentCut Node**: New node for cutting latents at precise points, enabling more granular control over latent space manipulation in complex generation workflows. This is particularly useful for batch processing and temporal video workflows.
19+
- **LatentCut Node**: New node for cutting latents at precise points, enabling more granular control over latent space manipulation in complex generation workflows. This is particularly useful for batch processing and temporal video workflows, such as removing specific frames from videos.
2020

2121
**Wan2.2 5B Model Integration**
22-
- **Fun Control Model Support**: Added support for Wan2.2 5B fun control model, expanding creative control options for video generation with enhanced parameter manipulation.
23-
- **Fun Inpaint Model Support**: Integrated Wan2.2 5B fun inpaint model, providing advanced inpainting capabilities for video workflows with improved quality and consistency.
22+
- **Fun Control Model Support**: Added support for Wan2.2 5B fun control model.
23+
- **Fun Inpaint Model Support**: Integrated Wan2.2 5B fun inpaint model.
2424

25-
**Template System Updates**
26-
- **Template v0.1.70**: Updated template system with latest model integrations and workflow examples, ensuring users have access to optimized setups for new features.
25+
**Workflow Templates**
26+
- **Template v0.1.70**: Added initial version of Wan2.2 S2V workflows
2727

28-
These improvements enhance ComfyUI's capabilities for audio-visual content creation, providing more reliable sound-to-video workflows and additional creative control through the latest Wan2.2 model series.
2928

3029
</Update>
3130

@@ -36,11 +35,10 @@ These improvements enhance ComfyUI's capabilities for audio-visual content creat
3635
This focused update improves the core node model patching system that underpins ComfyUI's flexible architecture:
3736

3837
**Core Infrastructure Enhancement**
39-
- **Node Model Patch Updates**: Enhanced nodes_model_patch.py with improvements to the underlying model patching mechanism that enables ComfyUI's dynamic node system and model loading capabilities
38+
- **Node Model Patch Updates**: Enhanced nodes_model_patch.py with improvements to the underlying model patching mechanism, making ComfyUI extensions for Qwen-Image ControlNet easier
4039

4140
**Workflow Benefits**
4241
- **Enhanced Stability**: Core model patching improvements contribute to more reliable node execution and model handling across different workflow configurations
43-
- **Foundation Update**: These backend improvements lay groundwork for better compatibility with future model architectures and node implementations
4442

4543
</Update>
4644

@@ -52,11 +50,11 @@ This release adds ComfyUI audio processing capabilities and includes performance
5250

5351
**Audio Processing Updates**
5452
- **Wav2vec2 Audio Encoder**: Added native wav2vec2 implementation as an audio encoder model, enabling audio-to-embedding workflows for multimodal applications
55-
- **Audio Encoders Directory**: Introduced dedicated models/audio_encoders directory structure for organized audio model management
53+
- **Audio Encoders Directory**: Added models/audio_encoders directory, which is the audio encoder directory for Wan2.2 S2V
5654
- **AudioEncoderOutput V3 Support**: Made AudioEncoderOutput compatible with V3 node schema, ensuring seamless integration with modern workflow architectures
5755

5856
**Google Gemini API Integration**
59-
- **Gemini Image API Node**: New Google Gemini Image API node by @Kosinkadink provides direct integration with Google's vision AI capabilities, expanding cloud-based image analysis options for workflows
57+
- **Gemini Image API Node**: Added new Google Gemini Image API node, the "nano-Nano-banana" image editing model API with high consistency
6058

6159
**Video Generation Performance & Memory Optimizations**
6260
- **WAN 2.2 S2V Model Support**: Work-in-progress implementation of WAN 2.2 Sound-to-Video model with optimized memory usage and performance
@@ -74,10 +72,9 @@ This release adds ComfyUI audio processing capabilities and includes performance
7472
- **Flux Memory Optimization**: Adjusted Flux model memory usage factors for better resource utilization
7573

7674
**Infrastructure & Reliability**
77-
- **Template Updates**: Updated to versions 0.1.66 and 0.1.68 with latest workflow templates and improvements
75+
- **Template Updates**: Updated to versions 0.1.66 and 0.1.68
7876
- **Documentation Cleanup**: Removed incompletely implemented models from readme to avoid user confusion
7977

80-
This release particularly benefits creators working with multimodal AI workflows combining audio, image, and video generation, while providing essential stability improvements for all users.
8178

8279
</Update>
8380

zh-CN/changelog/index.mdx

Lines changed: 13 additions & 18 deletions
Original file line numberDiff line numberDiff line change
@@ -6,26 +6,24 @@ icon: "clock-rotate-left"
66

77
<Update label="v0.3.55" description="2025年8月29日">
88

9-
**音频转视频工作流增强与模型支持扩展**
9+
**Wan2.2 S2V 工作流增强与模型支持扩展**
1010

11-
本次发布专注于改进音视频工作流功能并扩展模型支持,特别是针对音频转视频生成和最新的Wan2.2 5B模型
11+
本次发布专注于 Wan2.2 S2V 的相关视频工作流功能并扩展模型支持
1212

13-
**音频转视频工作流控制**
14-
- **WanSoundImageToVideoExtend节点**全新的音频转视频工作流手动视频扩展节点,为创作者提供对生成视频长度和时间的精确控制。这使得对音频内容如何转换为视频序列的精细调控成为可能。
13+
**Wan2.2 S2V工作流控制**
14+
- **WanSoundImageToVideoExtend节点**全新的音频驱动视频工作流手动视频扩展节点,为创作者提供对生成视频长度和时间的精确控制。这使得对音频内容如何转换为视频序列的精细调控成为可能。
1515
- **音视频同步**:修复了当视频扩展超过音频长度时导致工作流失败的关键问题,确保无论音频持续时间如何都能可靠地生成音频转视频内容。
1616
- **自动音频裁剪**:视频保存现在会自动将音频裁剪至匹配视频长度,消除最终输出文件中的音视频同步问题。
1717

1818
**高级潜在空间处理**
19-
- **LatentCut节点**:用于在精确点位切割潜在变量的新节点,在复杂生成工作流中实现对潜在空间操作的更精细控制。这对于批处理和时序视频工作流特别有用。
19+
- **LatentCut节点**:用于在精确点位切割潜在变量的新节点,在复杂生成工作流中实现对潜在空间操作的更精细控制。这对于批处理和时序视频工作流特别有用,比如删除视频中的某些帧
2020

2121
**Wan2.2 5B模型集成**
22-
- **Fun Control模型支持**:增加了对Wan2.2 5B fun control模型的支持,通过增强的参数操作为视频生成扩展了创意控制选项
23-
- **Fun Inpaint模型支持**:集成了Wan2.2 5B fun inpaint模型,为视频工作流提供高级修复功能,具有更好的质量和一致性
22+
- **Fun Control模型支持**:增加了对Wan2.2 5B fun control模型的支持,。
23+
- **Fun Inpaint模型支持**:集成了Wan2.2 5B fun inpaint模型。
2424

25-
**模板系统更新**
26-
- **模板v0.1.70**:更新了模板系统,集成了最新的模型和工作流示例,确保用户能够访问针对新功能优化的设置。
27-
28-
这些改进增强了ComfyUI在音视频内容创作方面的能力,通过最新的Wan2.2模型系列提供了更可靠的音频转视频工作流和额外的创意控制。
25+
**工作流模板**
26+
- **模板v0.1.70**:添加了 Wan2.2 S2V 初步版本的工作流
2927

3028
</Update>
3129

@@ -36,11 +34,10 @@ icon: "clock-rotate-left"
3634
此次专项更新改进了支撑 ComfyUI 灵活架构的核心节点模型补丁系统:
3735

3836
**核心基础设施增强**
39-
- **节点模型补丁更新**:增强了 nodes_model_patch.py,改进了底层模型补丁机制,该机制支持 ComfyUI 的动态节点系统和模型加载功能
37+
- **节点模型补丁更新**:增强了 nodes_model_patch.py,改进了底层模型补丁机制,使得 ComfyUI 对 Qwen-Image ControlNet 的扩展更为容易
4038

4139
**工作流程优势**
4240
- **增强的稳定性**:核心模型补丁改进有助于在不同工作流程配置中实现更可靠的节点执行和模型处理
43-
- **基础更新**:这些后端改进为未来模型架构和节点实现的更好兼容性奠定了基础
4441

4542
</Update>
4643

@@ -52,11 +49,11 @@ icon: "clock-rotate-left"
5249

5350
**音频处理更新**
5451
- **Wav2vec2 音频编码器**:新增原生 wav2vec2 实现作为音频编码器模型,为多模态应用启用音频到嵌入工作流
55-
- **音频编码器目录**引入专用的 models/audio_encoders 目录结构,用于有序的音频模型管理
52+
- **音频编码器目录**新增 models/audio_encoders 目录,此为 Wan2.2 S2V 音频编码器目录
5653
- **AudioEncoderOutput V3 支持**:使 AudioEncoderOutput 与 V3 节点架构兼容,确保与现代工作流架构的无缝集成
5754

5855
**Google Gemini API 集成**
59-
- **Gemini Image API 节点**@Kosinkadink 提供的全新 Google Gemini Image API 节点,可直接集成 Google 的视觉 AI 功能,为工作流扩展基于云的图像分析选项
56+
- **Gemini Image API 节点** 新增全新 Google Gemini Image API 节点,即 “nano-Nano-banana” 图像编辑模型 API,具有高度的一致性
6057

6158
**视频生成性能与内存优化**
6259
- **WAN 2.2 S2V 模型支持**:正在开发的 WAN 2.2 声音到视频模型实现,具有优化的内存使用和性能
@@ -74,11 +71,9 @@ icon: "clock-rotate-left"
7471
- **Flux 内存优化**:调整 Flux 模型内存使用因子以获得更好的资源利用
7572

7673
**基础设施与可靠性**
77-
- **模板更新**:更新至 0.1.66 和 0.1.68 版本,包含最新的工作流模板和改进
74+
- **模板更新**:更新至 0.1.66 和 0.1.68 版本
7875
- **文档清理**:从自述文件中移除未完全实现的模型,以避免用户混淆
7976

80-
此版本特别有利于从事结合音频、图像和视频生成的多模态 AI 工作流的创作者,同时为所有用户提供基本的稳定性改进。
81-
8277
</Update>
8378
<Update label="v0.3.52" description="2025年8月23日">
8479

0 commit comments

Comments
 (0)