r/StableDiffusion – Telegram
This media is not supported in your browser
VIEW IN TELEGRAM
Soprano 1.1-80M released: 95% fewer hallucinations and 63% preference rate over Soprano-80M

https://redd.it/1qcuuet
@rStableDiffusion
Media is too big
VIEW IN TELEGRAM
LTX-2 I2V synced to an MP3: Distill Lora Quality STR 1 vs .6 - New Workflow Version 2.

https://redd.it/1qd525f
@rStableDiffusion
Media is too big
VIEW IN TELEGRAM
I built a real-time 360 volumetric environment generator running entirely locally. Uses SD.cpp, Depth Anything V2, and LaMa, all within Unity Engine.

https://redd.it/1qde674
@rStableDiffusion
LTX-2 Updates

https://reddit.com/link/1qdug07/video/a4qt2wjulkdg1/player

We were overwhelmed by the community response to LTX-2 last week. From the moment we released, this community jumped in and started creating configuration tweaks, sharing workflows, and posting optimizations here, on, Discord, Civitai, and elsewhere. We've honestly lost track of how many custom LoRAs have been shared. And we're only two weeks in.

We committed to continuously improving the model based on what we learn, and today we pushed an update to GitHub to address some issues that surfaced right after launch.

What's new today:

Latent normalization node for ComfyUI workflows \- This will dramatically improve audio/video quality by fixing overbaking and audio clipping issues.

Updated VAE for distilled checkpoints \- We accidentally shipped an older VAE with the distilled checkpoints. That's fixed now, and results should look much crisper and more realistic.

Training optimization \- We’ve added a low-VRAM training configuration with memory optimizations across the entire training pipeline that significantly reduce hardware requirements for LoRA training. 

This is just the beginning. As our co-founder and CEO mentioned in last week's AMA, LTX-2.5 is already in active development. We're building a new latent space with better properties for preserving spatial and temporal details, plus a lot more we'll share soon. Stay tuned.



https://redd.it/1qdug07
@rStableDiffusion
This media is not supported in your browser
VIEW IN TELEGRAM
Viking Influencer with LTX2 image 2 video with native generated audio. It can do accents

https://redd.it/1qdu01e
@rStableDiffusion