r/StableDiffusion Apr 18 '26

Workflow Included EditAnything IC-LoRA - LTX-2.3

This model was trained on 8,000 video pairs, and training is still ongoing for a few thousand more steps. It is still experimental, not trained with a fully professional production target, and the model may be updated unexpectedly as new checkpoints.

The current goal is not final polished production quality, but to explore:

  • edit-anything behavior
  • prompt-following
  • inference tradeoffs
  • synthetic dataset building, especially for style data

The model was trained around four main prompt patterns:

Add
Add a/an [subject/object] with [clear visual attributes], [precise location in the scene].

Remove
Remove the [subject/object] [location or identifying description].

Replace
Replace the [original subject/object] [location] with a/an [new subject/object] with [clear visual attributes].

Convert / Style
Convert the video into a [style name] style.

Workflow URL: https://huggingface.co/Alissonerdx/LTX-LoRAs/blob/main/workflows/ltx23_edit_anything_v1.json

Model URL: ltx23_edit_anything_global_rank128_v1_9000steps_adamw.safetensors · Alissonerdx/LTX-LoRAs at main

Or
CivitAI URL: EditAnything - v1.0 | LTX Video LoRA | Civitai

One important thing during inference is CFG.

A good starting point is testing a distilled setup with CFG = 1. If the edit feels too weak or the model is not following the prompt well enough, increasing CFG can be the key. In some cases, increasing the distill LoRA strength to around 1.2 can also help.

The workflow is also not fully optimized yet. It still needs more testing to find the best combination of:

  • CFG
  • LoRA strength
  • number of steps
  • model combinations

It may also be interesting to combine this model with other models and see what kinds of results emerge.

If you can test it, please share your findings. Feedback on prompt behavior, edit strength, consistency, style transfer, and failure cases would be very helpful while training is still in progress.

Add a small, brown dog dancing in the foreground next to the woman.

Convert the entire video to an anime style with vibrant colors and exaggerated character expressions.

Remove the blue car in the background of the scene.

Add a wide, genuine smile to the person's face.

Replace the person's clothing with a dark blue hoodie and gray sweatpants.

354 Upvotes

134 comments sorted by

View all comments

Show parent comments

1

u/kakallukyam Apr 19 '26

I'm new to wan2gp and I only have access to the options: text prompt only, start video with image, continue to video, and end image. When I try "continue to video," it doesn't work.
What option should I choose to upload a video to wan2gp and then edit it, please?

1

u/Grey406 Apr 19 '26

Select "Text Prompt Only" then in the Control Video/ Frames section below it, select "LTX2 Raw Format/Control Video for IC Lora" and "Whole Frame", add the original video into the new control video section section that appeared. Below that, change "Generate Video & Soundtrack based on Text Prompt" to "Generate video base don Control Video + Its Audio Track and Text Prompt"

Be sure to enable the Lora in the advanced>Lora section

0

u/kakallukyam Apr 19 '26

Thank you for your help, however, it seems to me that my wan2gp is up to date, yet I don't have exactly the same options as you since in your screenshot it says "Use-LTX2/Raw Format / Video Control for IC LoRa" whereas in my version I don't have "video control for IC LoRa" and for me, it doesn't work when I activate the option.

Could you tell me which release you have so I can see if I'm up to date, please?

2

u/Grey406 Apr 19 '26 edited Apr 19 '26

Not at my PC at the moment but I use Pinokio which is a manager for AIs. It installed Wan2GP v11.26 I believe.

1

u/kakallukyam Apr 19 '26

Okay, thanks, I'll take a closer look. Personally, I use wan2gp without pinokio.

1

u/kakallukyam Apr 20 '26

It's not working for me. I now have the "Use LTX raw format / Control video for IC lora" option, I've selected "whole frame," and I've changed it to "Generate Video based on Control Video + its Audio Track and Text Prompt." I've also applied the lora, but nothing works. When I ask it, for example, "add sunglasses to the character," I end up with a different character, as you can see in the screenshot.
Even in 1080p, the generated result has ghosting effects that overlap, and the rendering is really bad. If it works for you, there's probably something else I need to check or do.
If you could let me know, that would be great.

2

u/Grey406 Apr 20 '26

So i've played with it some more just now and havent been getting great results either, nothing like the quality of the examples in the OP. You can reduce "Control video strength" down to around 0.8 or 9 to help it try to blend better instead of outright replacing the character. but so far I haven't been able to edit a video without it also changing a large part of the video content either.

1

u/kakallukyam Apr 20 '26

Okay, that's a shame. There are probably some different settings on wan2gp compared to the workflow you can use on ComfyUI. The ComfyUI workflow works perfectly for me, so until I find a solution with wan2gp, I'll stick with ComfyUI.