模型

LTX-Best-Face-ID

Alissonerdx/LTX-Best-Face-ID

查看上游原文 ↗
上游访问:公开 本站服务:可咨询 许可证:other 上游版本:dac8cc2dd6e3

中文简介

LTX-Best-Face-ID 是面向 LTX-2.3 的身份保持型参考图转视频 LoRA。上游说明它通过参考照片与文本提示生成尽量保持人物身份的视频,并提供近景人脸和角色设定图两种参考方式;使用真实人物素材时必须确认肖像、隐私和内容授权。

UPSTREAM README

上游模型卡 / 数据集卡

在 Hugging Face 查看原文 ↗

LTX-Best-Face-ID 是面向 LTX-2.3 的身份保持型参考图转视频 LoRA。上游说明它通过参考照片与文本提示生成尽量保持人物身份的视频,并提供近景人脸和角色设定图两种参考方式;使用真实人物素材时必须确认肖像、隐私和内容授权。

已有简体中文译文 · 本站中文整理 · 2026-07-23 14:50

LTX-Best-Face-ID — LTX-2.3 Identity LoRA (Reference-to-Video / IPT2V) An identity-preserving **reference-to-video** LoRA for **LTX-2.3 (22B)**. Give it a reference photo of a person + a text prompt, and it generates a video that keeps that person's identity. Built with **overlap reference conditioning + TASS-RoPE (source-phase RoPE)** and a differentiable ArcFace identity loss**. Runs in ComfyUI via the companion **BFS Nodes**. **Status:** two reference modes are now available from the same recipe: the original **close-up face** reference, and a newer **character-sheet** reference (face + full-body views) that also carries clothing/appearance consistency — see Character-Sheet Reference below. Other experimental variants (native Gemma-vision conditioning, timestep-split texture injection) may be released later if they prove out. 🎬 Examples 🎬 Face ID Base tag renders inline; .gif works too. Suggested layout: reference image (left) → generated video (right), with the prompt underneath. 🎬 Character Sheet What it does **Reference-to-video (ref_t2v):** one reference image → video of that identity performing the prompt's action. Identity is injected by placing the **reference latent** in the target's frame-0 RoPE grid (overlap) and tagging it with a distinct **source phase** so the model knows it is a *reference*, not the first frame to generate. An auxiliary **ArcFace face-similarity loss** on the decoded prediction sharpens the identity. How it works (technique) Overlap reference + TASS-RoPE (source-phase) The reference latent is concatenated to the video sequence sharing the frame-0 grid (classic IC-LoRA "overlap"). To stop the reference from leaking into / being confused with the generated first frame, each source gets a distinct **multiplicative RoPE phase**: This "source tag" lets the model separate *who is who* in the sequence and strongly improves identity transfer. Because the tag is positional, the same mechanism generalizes to **multiple references** (source_id

公开页仅展示原文摘录;完整模型卡或数据集卡请前往上游仓库查看。

上游文件元数据

  • .gitattributes4.48 KB
  • Best_FaceID_CharacterSheet_v1.0_LoRA.safetensors1.22 GB
  • Best_FaceID_v1.0_ArcFace_Projector.safetensors66.10 MB
  • Best_FaceID_v1.0_LoRA.safetensors2.30 GB
  • character_sheet_prompt.md4.22 KB
  • examples/references/celebs/Xtw4TNHkhCM_4_seg1_sheet.jpg330.91 KB
  • examples/references/celebs/Y9QI_u0OLBw_2_seg0_sheet.jpg430.84 KB
  • examples/references/celebs/YiIl9Uc-UOQ_17_seg2_sheet.jpg319.45 KB
  • examples/references/celebs/YiIl9Uc-UOQ_22_seg0_sheet.jpg337.78 KB
  • examples/references/celebs/YKx9UoT7JmE_3_seg0_sheet.jpg340.87 KB
  • examples/references/celebs/Ym3EvqlsfaA_15_seg0_sheet.jpg335.77 KB
  • examples/references/celebs/YmSABCyLEV0_2_seg0_sheet.jpg346.85 KB
第三方资源声明

本页面为橙子AI科技的中文整理与服务说明,不代表资源作者或平台官方页面。实际许可、访问和使用条件以上游原文为准。