Lightricks/LTX-2.3-22b-IC-LoRA-LipDub

Hugging Face Models Trending 模型

摘要

这个Hugging Face模型页面介绍了一个基于LTX-2.3-22b训练的IC-LoRA,用于唇语同步配音,包含项目页面、论文和推理流程。

Task: any-to-any Tags: diffusers, ltx-video, video-to-video, dubbing, lipdub, ic-lora, any-to-any, en, arxiv:2601.03233, arxiv:2601.22143, base_model:Lightricks/LTX-2.3, base_model:finetune:Lightricks/LTX-2.3, license:other, region:us
查看原文
查看缓存全文

缓存时间: 2026/05/17 18:27

Lightricks/LTX-2.3-22b-IC-LoRA-LipDub · Hugging Face

基于LTX-2.3-22b训练的IC-LoRA模型,用于在LTX模型上进行唇形配音。

该模型基于LTX-2(https://huggingface.co/papers/2601.03233)基础模型。

  • 项目页面: JustDubIt 项目页面 (https://justdubit.github.io/)
  • 论文: JustDubIt 论文 (https://arxiv.org/abs/2601.22143)
  • 代码: GitHub 仓库 (https://github.com/Lightricks/LTX-2)
  • 推理管线: packages/ltx-pipelines/src/ltx_pipelines/lipdub.py (https://github.com/Lightricks/LTX-2/blob/main/packages/ltx-pipelines/src/ltx_pipelines/lipdub.py)

模型文件

ltx-2.3-22b-ic-lora-lipdub-0.9.safetensors

许可

完整条款请参见LTX-2-社区许可

模型详情

  • 基础模型: LTX-2.3
  • 训练类型: IC-LoRA
  • 控制类型: 视频与音频
  • 参考下采样因子: 1(参考分辨率与输出分辨率相同)

🔌 在 ComfyUI 中使用

  1. 将 LoRA 权重复制到 models/loras 目录下。
  2. 使用 LTX-2 ComfyUI 仓库 (https://github.com/Lightricks/ComfyUI-LTXVideo/) 提供的官方唇形配音工作流。

数据集

该模型使用唇形配音数据集进行训练。

引用

@article{chen2026just,
  title={JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion},
  author={Chen, Anthony and Korem, Naomi Ken and Zeevi, Gal and Halperin, Tavi and Yosef, Matan Ben and Jelercic, Urska and Bibi, Ofir and Patashnik, Or and Cohen-Or, Daniel},
  journal={arXiv preprint arXiv:2601.22143},
  year={2026}
}

致谢

  • 基础模型由 Lightricks 提供
  • 训练基础设施:LTX-2 社区训练器

相似文章

Lightricks/LTX-2.3

Hugging Face Models Trending

Lightricks 发布了 LTX-2.3,这是一个基于扩散的开放权重音视频基础模型,具有改进的质量和提示遵循性,提供多个检查点,包括蒸馏和 LoRA 变体,可在本地执行。

Cseti/LTX2.3-22B_IC-LoRA-CrossView-Prompt

Hugging Face Models Trending

一个用于LTX-Video 2.3的概念验证型上下文LoRA适配器,能够通过固定词汇提示从新相机视角重新渲染视频场景,基于合成多视角数据训练而成。

LiconStudio/Ltx2.3-VBVR-lora-I2V

Hugging Face Models Trending

LiconStudio 发布了一个针对 LTX-2.3 的 LoRA 适配器,该适配器在 VBVR 数据集上进行了微调,以增强视频生成能力,改善提示理解、运动动态和时间一致性,用于复杂的视频推理任务。

Lightricks/LTX-2.5

Hugging Face Models Trending

Lightricks releases LTX-2.5, an open-weights world model for generating synchronized video and audio from text, image, and video inputs, with features like native multishot generation and a new diffusion video decoder.

Lightricks/LTX-2

GitHub Trending (daily)

LTX-2 是 Lightricks 推出的首个基于 DiT 的音频-视频基础模型,提供同步音频和视频生成、高保真度以及可投入生产的输出,并附带开源代码和开放模型权重。