JoaoZaokk/LTX-2.5-22B-distilled-W4A8-ConvRot
text-to-video model by JoaoZaokk
From the publisher
Model card & documentation
Source previewRead the publisher’s intended use, setup instructions, evaluations and limitations. The original model card is the source of truth.
A asymw4a8int8 build of ltx-2.5-22b-distilled-transformer, measured on a complete 10-second video and its soundtrack against the BF16 original and against Lightricks' own INT8. 39.13 GiB → 11.66 GiB, 3.36x lighter, and 1.95x faster than the original on the same 249-frame render. It is also 1.90x less faithful in the picture and 2.9x less faithful in the audio than Lightricks' own INT8 build — those numbers are here because they are the ones that decide whether you want this file. Method, tools and the full measurement log: https://github.com/JoaoZaokk/comfy-quant-bench LTX 2.5 generates video and audio in one latent. The first version of this card decoded only the video branch, and said so in its "not covered" list as if that were a footnote. It is not: an audio-video model judged on frames alone is judged on half its output. All three arms were re-rendered with the audio branch decoded (LTXVAudioVAEDecode on the…
Read the full model card ↗ · Preview checked 2026-09-28T06:39:06.033Z
Inside the original model card — Document outline
- LTX 2.5 22B distilled — 4-bit weights, a real 10-second video with its audio, and the arm that beats
- Correction, 2026-09-13: the first version of this card measured half of what the model makes
- The measurement: one 10-second video, three transformers, picture and sound
- Picture
- Sound
- Two more int8 builds, added 2026-09-14: the rotation is the recipe
- Listen for yourself
- The file
- Running it
- What is NOT covered
- License and changes
- Credits
Headings are captured from the source. Links open the publisher’s document, not a locally hosted copy.
Documentation belongs to its respective authors. Reported project/model license: other. A listing is not a grant of reuse or training rights. Confirm the document’s own terms at the source.
Model overview
JoaoZaokk/LTX-2.5-22B-distilled-W4A8-ConvRot is a text-to-video model repository published by JoaoZaokk on Hugging Face. The source reports the diffusion-single-file library.
This page summarizes Hub metadata. For intended use, training data, evaluation results and limitations, consult the original model card.
Model facts
- Task
- text-to-video
- Library
- diffusion-single-file
- Recent downloads (30 days)
- 43,168
- Cumulative likes
- 0
- Hugging Face trending score
- 0
- Reported safetensors parameters
- Not reported
- Architecture
- Not reported
- License
- other
- Access
- Not gated by Hugging Face
- Created
- 2026-09-13T03:16:36.000Z
- Last modified
- 2026-09-14T07:48:42.000Z
Compatibility and lineage
Reported languages: en
Reported base models: Lightricks/LTX-2.5
Parameter count is not a RAM/VRAM requirement. Check precision, quantization, context length and runtime compatibility in the model card; no hardware or API-cost claim is inferred here.
Use and evaluate
Check license terms and access requirements first. Review the model card’s documented loading instructions, evaluations and safety limitations. Benchmark scores are not imported by this directory, and popularity is not an accuracy ranking.
No model weights are downloaded or executed by AltAPIs. Never enable remote model code without reviewing it.
Source and freshness
Source: Hugging Face Hub. Metadata observed 2026-09-28T06:38:59.518Z. Daily imports are snapshots, not real-time monitoring.
Popularity and source listings do not establish security, suitability, licensing rights or benchmark performance.
Related models
FastVideo/FastVideo-FastH3-4-step-Preview-v1-VSA-DataFree
QuantStack/Wan2.2-T2V-A14B-GGUF
Wan-AI/Wan2.1-T2V-1.3B-Diffusers
larryvrh/MiniMax-H3-Turbo-Lora