跳到正文
@omarsar0· @omarsar0 · X·· 27 天前AI 评分61
AI 导读

LTX-2.5 是 ltx_io 推出的开放权重视频世界模型,支持本地部署,并提供可微调的预训练基座,作者提到上一代 LTX 下载量已达 18M。该版本同时改进生成与编辑,新解码器针对更清晰的人脸、可读的文字与招牌以及更干净的快速运动,Native Multishot 能生成在角色、环境、光照和声音上保持连贯的多镜头。

正文

Open video models are having their moment.

The previous generation of LTX alone reached 18M downloads, which shows how much demand there is for video models that builders can own and modify themselves.

LTX-2.5 from @ltx_io builds on that foundation as a world model with open weights, local deployment, and a pretrained base that teams can fine-tune for their own work.

For me, it's all about owning the intelligence stack.

Builders choose the hardware, keep access to the weights, and can fine-tune the model around a creative or production workflow.

More capabilities related to this model:

The release improves both generation and editing.

The new decoder targets sharper faces, legible text and signage, and cleaner fast motion.

Native Multishot generates connected shots while preserving character, environment, lighting, and voice across cuts.

IC-LoRA works on footage you already have, with support for object removal, continuity and wardrobe fixes, and environment changes without a reshoot or frame-by-frame rotoscoping.

A stronger distilled model brings more of the full model's quality and motion to local GPUs.

Open video is still early, and releases like this give builders more room to experiment, specialize models, and own the production stack.

I am looking forward to seeing what people build with LTX-2.5.

来源:@omarsar0 · x.com