CoolFace
Modelpublic

SyFeee/LTX-2.3-SyFe-Plain-AV-LoRA

sourceHugging Faceotherupdated 1mo agoView on Hugging Face
1likes74downloads
Model Card

SyFe LTX-2.3 Plain and Audio-Video LoRAs

Self-trained SyFe LoRA checkpoints for Chinese-drama generation, bilingual prompting, and joint audio-video experiments on LTX-2.3 22B-dev.

Checkpoints

RunPurposeRankStepsStatus
tv306954_run01First multi-character show baseline643,000Kept baseline; final loss 0.2897
v5_unifiedRich 747-clip Chinese-drama corpus324,000Shipped plain LoRA
official_704_cuvalNative 1280x704 official-trainer control run323,000Experimental control
SyFe_Bilingual_Plain_704Bilingual 704p production stack component3210,000Deployed checkpoint
av_lora_v1_productionJoint audio-video experiment644,000Archived
av_lora_v2_productionJoint audio-video P10.5 experiment644,000Kept with caveats

The AV checkpoints can generate ambient music, breathing, and sound effects, but they did not produce reliable intelligible dialogue. They should not be described as standalone voice-cloning models; use the ID-LoRA release for reference-voice conditioning.

Run folders include final weights and available training configs. v5_unified and the bilingual deployment checkpoint are release copies whose original preprocessing data is not included.

Use is subject to the LTX-2 community license.