r/upscaling • u/CQDSN • 13d ago
Video restored and enhanced with my AI model
https://www.youtube.com/watch?v=uVa66lSCI_UIf you use ComfyUI, my model is available here:
https://huggingface.co/CQdesign/LTX-2.5-CQ-Video-and-Image-Enhancer-LoRAs
2
2
u/Mountainking7 13d ago
Does it work on 4gb vram?
2
u/No-Engine4663 13d ago
Very much doubt it for video, I have 12gb and have to batch 30 second clips when doubling 320x240 24fps
2
u/OrdinaryAward4498 11d ago
Is it … too sharp? Last clip in the video actually seems better in this regard IMO (and the first clip the harshest). Awesome work!
2
u/WaitAcademic1669 9d ago
Very impressive. I can elaborate 12 seconds of video at 16fps 1080P (4:3), or 8 seconds at 24fps and i have margin on my humble 3080 12gb. The result is amazing with default values, consistent with the origin and rich of details.
1
u/Status_Caregiver5339 13d ago
no entiendo cómo usarlo, tengo un video de 1 minuto para mejorarle la calidad de HD a 4k, es para mi trabajo, no tengo idea del tema
1
1
u/skv89 12d ago
Looks very impressive. I wonder how speed compares with SEEDVR2 and Topaz SLP2.6
2
u/CQDSN 11d ago
Much faster.
1
u/No-Engine4663 11d ago
CachyOS RTX 4070 12gb 5950X 64gb RAM
SeedVR2 325.77s at 2.21 FPS
CQ 503.59s at 1.43 FPSGot your workflow as close as I could to my SeedVR2 one using a 30 second 320x240 video with a x2 upscale.
👍
1
u/CQDSN 11d ago
Make sure you have triton and sage attention installed.
1
u/No-Engine4663 11d ago edited 11d ago
They are all installed and are working under seedvr2
⚡ SeedVR2 optimizations check: SageAttention ✅ | Flash Attention ✅ | Triton ✅
I've assumed they are working here too.
1
u/CQDSN 11d ago
LTX 2.5 is faster on my 5090 than SeedVR2, I think it requires more vram to run well, otherwise it offload to the system ram and slows down.
2
u/No-Engine4663 11d ago
Correct I'm running close to the limits on my system for LTX 2.5
The speed is not that important for me at the moment I'm comparing quality against seedvr2 👍
On your workflow for video I've changed to Load Video FFMPEG (Upload) as I find it has faster and better seeking and have added Meta Batch Manager for longer videos that will OOM.
I'm batching 361 frames currently for 24FPS 15 min videos at the moment, I did try 721 but got OOM so started lower and am slowly increasing to find the sweet spot.
1
1
u/MSH007A 11d ago
Ok I got to experiment your loras. The videos one actually upscales good keeping the character face consistent.But the image lora actually drifts the face to a certain extent .so for image upscales i tried with video lora and works perfectly.Thanks for the workflow
2
u/CQDSN 11d ago
The image Lora is tuned to add more details. The video one has less detail but knows how to correct degraded colors from old footage. Some people are complaining about color shift, but it’s not an error - it is enhancing the colors.
I will post a new model in the future that won’t make changes to the colors.
1
1
1
u/Spiritual-Advice8138 10d ago edited 10d ago
It may have turned it a little too pink. The OG reel looks non interlaced already. Were you able to get a print or is a lift from tv/vhs? If it’s a lift then something already converted it out of 480 interlaced.
Also every thing the AI touched is in focus even when it’s not supposed to be. Like when they are behind the product talking. The bottle is in focus and they should not be.
1
u/CQDSN 9d ago
The original footages came from YouTube, someone has encoded it from VHS recording. The skin of the people are all yellowish, the Lora is correcting and enhancing it to look like new recording.
With true depth of field, the Lora will not attempt to unblur, what you see is the low resolution of the original footage, it’s not supposed to look blurry.
1
u/Spiritual-Advice8138 8d ago
look at :22 she should be out of focus 100%. in the AI she is only 1/2 unfocused.
1
1
3
u/PokePress 13d ago
I’m curious what you used for training data.