Nothing here is a floating tag. The recipe checks each of these before it serves a token.
- Target checkpoint
Mia-AiLab/GLM-5.3-Flash-EXL3-TR3-4bpw at revision 25a44fdbf16862a46b7cc9921142c6c81350af2f, itself byte-identical to brandonmusic/GLM-5.3-Flash-tr3-4bpw at 5ab363a8dcf6405955fd5f99671e01a1c9fb124b. The JSpark3 Hugging Face repository re-hosts this revision shard for shard with the same hashes; the preflight accepts either source because the bytes are identical- Draft checkpoint
incoai/GLM-5.3-Flash-DFlash2 at revision dc77ff1c99eeb2df044ee3d4f0094eb033fee410, k=7- Container image
ghcr.io/miaai-lab/glm-5.3-flash-2x-dgx-sparks at digest sha256:9bb1557a4234fce63d59599e44d10747eabd742beb337eebf9e7070be8a0fd58, launched by digest, not redistributed- Serving engine
- vLLM build
487ecf187 as shipped inside the pinned image; five hash-gated runtime transforms are applied at start - Transform sources
- FlyCockpit
GLM-5.3-Flash-EXL3-3x-DGX-Sparks at 9093765c757bd1976372196e44af84a67cf86bad; vcruz305 GLM-5.3-Flash-EXL3-K2-DGX-Spark-recipe at 622cb878d66f703c597bd6baaa2423caa1786f99 - Runtime envelope
- Configured context 1,000,000 tokens · 32 sequences · 8,192 batched tokens · GPU memory utilization 0.83 · FP8 KV cache · prefix caching · served as
glm-5.3-flash