ukisai/Swift-1.5-Qwen3.8-Flash-Next-GSQ-RCO-GGUF
image-text-to-text model by ukisai
From the publisher
Model card & documentation
Source previewRead the publisher’s intended use, setup instructions, evaluations and limitations. The original model card is the source of truth.
Website • Learn more • BF16 model • Standard GGUFs • Evaluation • Enterprise licensing Mixed-precision GGUF quantizations of Swift Flash Next, with Swift-specific GSQ refinement and reused ISTA GSQ-RCO per-tensor allocation profiles. Swift Flash Next is UkisAI's reasoning-efficient derivative of Qwen3.8-Flash-Next. Its post-training targets shorter reasoning traces and coding, agentic and long-horizon tasks. See the original model card for model-level benchmarks and training details. Those benchmarks are separate from the quantization measurements below. Swift 1.5 Flash-Next uses 63.4% fewer thinking tokens, with a 1.8x speed up while keeping the accuracy loss <1% vs base on xhigh. Each tier contains two GGUF shards. Download both files into the same directory and load shard 1; llama.cpp locates shard 2 automatically. Sizes are decimal GB and exclude runtime context/cache memory. Tier names describe mixed-precision allocation profiles. A BF16 vision projector is included separately (0.91 GB). The evaluation below covers text inference; it…
Read the full model card ↗ · Preview checked 2026-09-29T09:22:08.459Z
Inside the original model card — Document outline
- Swift 1.5 Qwen3.8-Flash-Next · GSQ-RCO
- Available quantizations
- Evaluation
- Usage
- Quantization procedure
- Methods and acknowledgements
- License and access
Headings are captured from the source. Links open the publisher’s document, not a locally hosted copy.
Links from the model card
References supplied by the publisher, not independently verified endorsements. Check the destination before downloading files or entering credentials.
Documentation belongs to its respective authors. Reported project/model license: other. A listing is not a grant of reuse or training rights. Confirm the document’s own terms at the source.
Model overview
ukisai/Swift-1.5-Qwen3.8-Flash-Next-GSQ-RCO-GGUF is a image-text-to-text model repository published by ukisai on Hugging Face. The source reports the gguf library.
This page summarizes Hub metadata. For intended use, training data, evaluation results and limitations, consult the original model card.
Model facts
- Task
- image-text-to-text
- Library
- gguf
- Recent downloads (30 days)
- 23,970
- Cumulative likes
- 54
- Hugging Face trending score
- 53
- Reported safetensors parameters
- Not reported
- Architecture
- Not reported
- License
- other
- Access
- Not gated by Hugging Face
- Created
- 2026-09-24T14:11:18.000Z
- Last modified
- 2026-09-24T16:16:42.000Z
Compatibility and lineage
Language coverage was not reported.
Reported base models: ukisai/Swift-Qwen3.8-Flash-Next
Parameter count is not a RAM/VRAM requirement. Check precision, quantization, context length and runtime compatibility in the model card; no hardware or API-cost claim is inferred here.
Use and evaluate
Check license terms and access requirements first. Review the model card’s documented loading instructions, evaluations and safety limitations. Benchmark scores are not imported by this directory, and popularity is not an accuracy ranking.
No model weights are downloaded or executed by AltAPIs. Never enable remote model code without reviewing it.
Source and freshness
Source: Hugging Face Hub. Metadata observed 2026-09-29T06:38:06.650Z. Daily imports are snapshots, not real-time monitoring.
Popularity and source listings do not establish security, suitability, licensing rights or benchmark performance.