中文简介
LTX-Best-Face-ID 是面向 LTX-2.3 的身份保持型参考图转视频 LoRA。上游说明它通过参考照片与文本提示生成尽量保持人物身份的视频,并提供近景人脸和角色设定图两种参考方式;使用真实人物素材时必须确认肖像、隐私和内容授权。
上游模型卡 / 数据集卡
LTX-Best-Face-ID 是面向 LTX-2.3 的身份保持型参考图转视频 LoRA。上游说明它通过参考照片与文本提示生成尽量保持人物身份的视频,并提供近景人脸和角色设定图两种参考方式;使用真实人物素材时必须确认肖像、隐私和内容授权。
LTX-Best-Face-ID — LTX-2.3 Identity LoRA (Reference-to-Video / IPT2V) An identity-preserving **reference-to-video** LoRA for **LTX-2.3 (22B)**. Give it a reference photo of a person + a text prompt, and it generates a video that keeps that person's identity. Built with **overlap reference conditioning + TASS-RoPE (source-phase RoPE)** and a differentiable ArcFace identity loss**. Runs in ComfyUI via the companion **BFS Nodes**. **Status:** two reference modes are now available from the same recipe: the original **close-up face** reference, and a newer **character-sheet** reference (face + full-body views) that also carries clothing/appearance consistency — see Character-Sheet Reference below. Other experimental variants (native Gemma-vision conditioning, timestep-split texture injection) may be released later if they prove out. 🎬 Examples 🎬 Face ID Base tag renders inline; .gif works too. Suggested layout: reference image (left) → generated video (right), with the prompt underneath. 🎬 Character Sheet What it does **Reference-to-video (ref_t2v):** one reference image → video of that identity performing the prompt's action. Identity is injected by placing the **reference latent** in the target's frame-0 RoPE grid (overlap) and tagging it with a distinct **source phase** so the model knows it is a *reference*, not the first frame to generate. An auxiliary **ArcFace face-similarity loss** on the decoded prediction sharpens the identity. How it works (technique) Overlap reference + TASS-RoPE (source-phase) The reference latent is concatenated to the video sequence sharing the frame-0 grid (classic IC-LoRA "overlap"). To stop the reference from leaking into / being confused with the generated first frame, each source gets a distinct **multiplicative RoPE phase**: This "source tag" lets the model separate *who is who* in the sequence and strongly improves identity transfer. Because the tag is positional, the same mechanism generalizes to **multiple references** (source_id
上游文件元数据
.gitattributes4.48 KBBest_FaceID_CharacterSheet_v1.0_LoRA.safetensors1.22 GBBest_FaceID_v1.0_ArcFace_Projector.safetensors66.10 MBBest_FaceID_v1.0_LoRA.safetensors2.30 GBcharacter_sheet_prompt.md4.22 KBexamples/references/celebs/Xtw4TNHkhCM_4_seg1_sheet.jpg330.91 KBexamples/references/celebs/Y9QI_u0OLBw_2_seg0_sheet.jpg430.84 KBexamples/references/celebs/YiIl9Uc-UOQ_17_seg2_sheet.jpg319.45 KBexamples/references/celebs/YiIl9Uc-UOQ_22_seg0_sheet.jpg337.78 KBexamples/references/celebs/YKx9UoT7JmE_3_seg0_sheet.jpg340.87 KBexamples/references/celebs/Ym3EvqlsfaA_15_seg0_sheet.jpg335.77 KBexamples/references/celebs/YmSABCyLEV0_2_seg0_sheet.jpg346.85 KB
本页面为橙子AI科技的中文整理与服务说明,不代表资源作者或平台官方页面。实际许可、访问和使用条件以上游原文为准。