Language Models


Name Total Parameters Quantization Context Length Available until More information
Kimi-K2.6 Total: 1.1T / Activated: 32B Native INT4 (QAT) 262,144 tokens 20/11/26 https://huggingface.co/moonshotai/Kimi-K2.6
GLM-5.3-Flash Total: 320B / Activated: 18B NVFP4 (weights only) 1,048,576 tokens 20/11/26 https://huggingface.co/LibertAIDAI/GLM-5.3-Flash-NVFP4
DeepSeek-V4-Flash Total: 284B / Activated: 13B Mixed FP4 + FP8 (experts in FP4) 262,144 tokens 20/11/26 https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-DSpark
Qwen3.8 Flash Next Total: ~180B / Activated: ~6B FP8 262,144 tokens 20/11/26 https://huggingface.co/Qwen/Qwen3.8-Flash-Next-FP8
Qwen3.6 35B-A3B Total: 35.1B / Activated: 3B BF16 262,144 tokens 20/11/26 https://huggingface.co/Qwen/Qwen3.6-35B-A3B
Qwen3 32B 32.8B (dense) BF16 30,000 tokens 20/11/26 https://huggingface.co/Qwen/Qwen3-32B

Public models are available to the entire SCAYLE user community with active permissions to use this service. Each model included in this category is offered with an initial availability cycle of 90 days. Once this period has elapsed, the SCAYLE team will review its usage, performance, and demand to decide whether it should be maintained, renewed, or replaced with another alternative. Users are advised to consult this section periodically to check the updated status of each model.

Private models are restricted-access resources that are not shown by default in the general catalogue. To access them, they must be explicitly requested from the SCAYLE team via or through a ticket in GLPI. Permission granting is a manual process assessed by the administrators, who will decide whether access is appropriate based on resource availability and the requirements of the project. If, after receiving access confirmation, you do not see the model in your Caléndula environment or in the web interface, please notify .