Views
No views yet
| Field | Value |
|---|---|
| Base model | Lightricks/LTX-2.3-22B |
| Adapter type | Plain LoRA (PEFT-style) |
| Rank | 64 |
| Alpha | 64 |
| Target modules | to_k, to_q, to_v, to_out.0 |
| Training steps | 3000 |
| Optimizer | AdamW |
| Learning rate | 1e-4, linear schedule |
| Mixed precision | bf16 |
| Gradient checkpointing | enabled |
char_0_person, char_1_person, ...) substituted at the start of each prompt, followed by camera/scene prose and a style anchor (live-action photorealistic, cinematic Chinese drama).char_0_person, char_1_person, etc. correspond to the most-frequent identity clusters discovered by ArcFace + DBSCAN over the corpus. They must appear at the start of the prompt, comma-separated, terminated with a period.char_0_person, char_1_person. Framed in a static eye level medium shot,
on a 35mm normal lens, with natural light. Set in a torch-lit Han dynasty
courtyard at dusk, the subjects face each other in tense silence.
Live-action photorealistic, cinematic Chinese drama.diffusers1from diffusers import LTXPipeline
2import torch
3
4pipe = LTXPipeline.from_pretrained(
5 "Lightricks/LTX-2.3-22B",
6 torch_dtype=torch.bfloat16,
7)
8pipe.to("cuda")
9
10# Load the LoRA
11pipe.load_lora_weights(
12 "SyFeee/ltx2.3-chinese-drama-charlora",
13 weight_name="lora_weights_step_03000.safetensors",
14 adapter_name="cn_drama_char",
15)
16pipe.set_adapters(["cn_drama_char"], adapter_weights=[0.9])
17
18video = pipe(
19 prompt=(
20 "char_0_person. Framed in a close-up on a 50mm normal lens, with shallow focus. "
21 "Set in a candlelit Han dynasty study, the subject sits writing on bamboo scrolls. "
22 "Live-action photorealistic, cinematic Chinese drama."
23 ),
24 negative_prompt="no CGI, no animation, no illustration, no painterly style, no anime",
25 width=1280,
26 height=544,
27 num_frames=89,
28 guidance_scale=4.0,
29 num_inference_steps=20,
30).frames[0]0.8 — subtle stylistic touch, identity present but soft0.9 — default, validated against training distribution1.0+ — risks overfitting on facial features at the expense of camera/scene freedomdolly in, 35mm normal lens, handheld).char_{id}_person. Cluster frequencies in the training corpus:| Trigger | Clips |
|---|---|
char_0_person | 442 |
char_0_person, char_1_person | 180 |
char_0_person, char_1_person, char_2_person | 53 |
| 4–10 character combinations | 49 |
| No characters detected (style-only) | 20 |
SyFeee/ltx2.3-chinese-drama-iclora-pose — pose-controlled IC-LoRA on the same corpus.SyFeee/ltx2.3-chinese-drama-iclora-depth — depth-controlled IC-LoRA.SyFeee/ltx2.3-chinese-drama-iclora-canny — canny-edge-controlled IC-LoRA.LICENSE for terms.