Text / dealignai
GLM 5.3 CYBERSECURITY FP8
GLM 5.3 CYBERSECURITY FP8 from dealignai. Source-based hardware guidance from its published configuration.
EstimatedRepository opened Aug 30, 2026Source checked 9/23/2026Version: 5915c1b8
LOCALRENTED GPUOPEN WEIGHTS
At a glance
- Parameters
- 753.33B
- Architecture
- glm_moe_dsa
- Context length
- 1,048,576
- License
- mit
- Software
- vLLM
Will it run on your machine?
Save your machine to see a personalized rating and its reasoning.
Add my hardwareWays to run it
vLLM · Linux
Repository-specific command in publisher documentation; confirm dependencies and hardware in the source.
vllm serve dealignai/GLM-5.3-CYBERSECURITY-FP8 --tensor-parallel-size 8 --gpu-memory-utilization 0.90 --enforce-eager --disable-custom-all-reduce --enable-prefix-caching --max-num-seqs 24 --max-model-len 131072 --reasoning-parser glm45 --tool-call-parser glm47 --enable-auto-tool-choiceExplore its uses
SAFETENSORSTEXT-GENERATION
Keep exploring
Text
Prism MLBonsai 2 · 27B
A compressed 27B-class reasoning model with publisher-provided GGUF packs and a dedicated llama.cpp fork for CUDA, Metal, and CPU.
MODEL SIZE27.36B
Text
DeepSeekDeepSeek V4.1 Flash
A multimodal reasoning model with a compressed key-value cache, published as open weights.
Text
QwenQwen3.8-Flash-Next
An experimental open-weight multimodal model with sparse attention and 262K native context.
MODEL SIZE125B language · 6B active