TheMindExpansionNetwork commited on
Commit
cddc3f6
·
verified ·
1 Parent(s): 386e324

Upload folder using huggingface_hub

Browse files
.gitattributes CHANGED
@@ -33,3 +33,5 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
 
 
 
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
36
+ samples/sample_0.gif filter=lfs diff=lfs merge=lfs -text
37
+ samples/sample_1.gif filter=lfs diff=lfs merge=lfs -text
README.md ADDED
@@ -0,0 +1,63 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ tags:
3
+ - ltx-2
4
+ - ltx-video
5
+ - text-to-video
6
+ - audio-video
7
+ pinned: true
8
+ language:
9
+ - en
10
+ license: other
11
+ pipeline_tag: text-to-video
12
+ library_name: diffusers
13
+ ---
14
+
15
+ # swing-u-sinnerz
16
+
17
+ This is a fine-tuned version of [`ltx-2-19b-dev.safetensors`](/mnt/ltx-trainer/modals/ltx-2-19b-dev/ltx-2-19b-dev/ltx-2-19b-dev.safetensors) trained on custom data.
18
+
19
+ ## Model Details
20
+
21
+ - **Base Model:** [`ltx-2-19b-dev.safetensors`](/mnt/ltx-trainer/modals/ltx-2-19b-dev/ltx-2-19b-dev/ltx-2-19b-dev.safetensors)
22
+ - **Training Type:** LoRA fine-tuning
23
+ - **Training Steps:** 2000
24
+ - **Learning Rate:** 0.0001
25
+ - **Batch Size:** 1
26
+
27
+ ## Sample Outputs
28
+
29
+ | | | | |
30
+ |:---:|:---:|:---:|:---:|
31
+ | ![example1](./samples/sample_0.gif)<br><details style="max-width: 300px; margin: auto;"><summary>Prompt</summary>A black-and-white 1930s rubber-hose cartoon character with stretchy limbs walks down a crooked city street at night under a flickering streetlamp, pausing to glance at a shadow that slowly crawls along a brick wall behind them. The characters eyes widen and their mouth stretches in a startled expression, then they shuffle backward with bouncy steps as the shadow grows larger. The audio includes a soft projector hum and film crackle, light tap-tap footsteps on pavement, a quiet spooky jazz riff from muted brass and upright bass, and a short cartoon whoosh as the shadow slides across the wall.</details> | ![example2](./samples/sample_1.gif)<br><details style="max-width: 300px; margin: auto;"><summary>Prompt</summary>A black-and-white 1930s rubber-hose cartoon character sits at a small upright piano in a surreal underworld room with warped picture frames and swaying candles, playing a simple jazzy rhythm while their head and shoulders bob in time. As the melody continues, the room’s shadows stretch and bend in odd shapes and the character looks left and right with an uneasy grin, then pulls their hands back as if something is moving inside the piano. The audio features soft piano notes with a light swing feel, subtle projector hiss and film crackle, gentle creaks from the room, and brief cartoon squeaks timed to the character’s movements.</details> |
32
+
33
+ ## Usage
34
+
35
+ This model is designed to be used with the LTX-2 (Lightricks Audio-Video) pipeline.
36
+
37
+ ### 🔌 Using Trained LoRAs in ComfyUI
38
+
39
+ In order to use the trained LoRA in ComfyUI, follow these steps:
40
+
41
+ 1. Copy your trained LoRA checkpoint (`.safetensors` file) to the `models/loras` folder in your ComfyUI installation.
42
+ 2. In your ComfyUI workflow:
43
+ - Add the "Load LoRA" node to choose your LoRA file
44
+ - Connect it to the "Load Checkpoint" node to apply the LoRA to the base model
45
+
46
+ You can find reference Text-to-Video (T2V) and Image-to-Video (I2V) workflows in the
47
+ official [LTX-2 repository](https://github.com/Lightricks/LTX-2).
48
+
49
+ ### Example Prompts
50
+
51
+ Example prompts used during validation:
52
+
53
+ - `A black-and-white 1930s rubber-hose cartoon character with stretchy limbs walks down a crooked city street at night under a flickering streetlamp, pausing to glance at a shadow that slowly crawls along a brick wall behind them. The characters eyes widen and their mouth stretches in a startled expression, then they shuffle backward with bouncy steps as the shadow grows larger. The audio includes a soft projector hum and film crackle, light tap-tap footsteps on pavement, a quiet spooky jazz riff from muted brass and upright bass, and a short cartoon whoosh as the shadow slides across the wall.`
54
+ - `A black-and-white 1930s rubber-hose cartoon character sits at a small upright piano in a surreal underworld room with warped picture frames and swaying candles, playing a simple jazzy rhythm while their head and shoulders bob in time. As the melody continues, the room’s shadows stretch and bend in odd shapes and the character looks left and right with an uneasy grin, then pulls their hands back as if something is moving inside the piano. The audio features soft piano notes with a light swing feel, subtle projector hiss and film crackle, gentle creaks from the room, and brief cartoon squeaks timed to the character’s movements.`
55
+
56
+
57
+
58
+ This model inherits the license of the base model ([`ltx-2-19b-dev.safetensors`](/mnt/ltx-trainer/modals/ltx-2-19b-dev/ltx-2-19b-dev/ltx-2-19b-dev.safetensors)).
59
+
60
+ ## Acknowledgments
61
+
62
+ - Base model: [Lightricks](https://huggingface.co/Lightricks/LTX-2)
63
+ - Trainer: [LTX-2](https://github.com/Lightricks/LTX-2)
lora_weights_step_02000.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:bde9a2742220012f3f72afda53c230eb3263280a36681e423088db11c9baa07a
3
+ size 428150664
samples/sample_0.gif ADDED

Git LFS Details

  • SHA256: 5695971ce2ff71794b3e0600cc23481cfb242441c99cceb7cd33eb91e6c5b4d6
  • Pointer size: 133 Bytes
  • Size of remote file: 21.1 MB
samples/sample_1.gif ADDED

Git LFS Details

  • SHA256: 3b931b74c5273abbe9000765a527fd37fbaa043a97d28a3cd421d6ceebc3be56
  • Pointer size: 133 Bytes
  • Size of remote file: 18.9 MB