Repository: localaiLicense: mit
Ornith-1.5-35B-A3B is an MIT-licensed Qwen3.5 mixture-of-experts model from Ornith AI for agentic coding, reasoning, repository-level software tasks, and tool use. It supports text and image input with a context window of 262K tokens. This default entry uses the APEX Balanced GGUF and BF16 vision projector. Compact APEX and MTP-enabled APEX builds are available as variants.
Links
Tags
Ornith-1.5-35B-A3B in the smaller APEX Compact GGUF format, with the shared BF16 vision projector for multimodal prompts.
Links
Tags
Ornith-1.5-35B-A3B in the APEX Balanced GGUF format with native multi-token prediction enabled for speculative decoding, plus the shared BF16 vision projector.
Links
Tags
Ornith-1.5-35B-A3B in the APEX Compact GGUF format with native multi-token prediction enabled for speculative decoding, plus the shared BF16 vision projector.
Links
Tags
Repository: localaiLicense: mit
Ornith-1.5-35B-A3B is an MIT-licensed Qwen3.5 mixture-of-experts model from Ornith AI for agentic coding, reasoning, repository-level software tasks, and tool use. It activates about 3B parameters per token and supports text and image input with a context window of 262K tokens. This default entry uses the Q4_K_M GGUF and BF16 vision projector. Q5_K_M, Q6_K, and Q8_0 builds are available as variants.
Links
Tags
Ornith-1.5-35B-A3B in the higher-quality Q8_0 GGUF format, with the shared BF16 vision projector for multimodal prompts.
Links
Tags
Ornith-1.5-35B-A3B in the Q5_K_M GGUF format, with the shared BF16 vision projector for multimodal prompts.
Links
Tags
Ornith-1.5-35B-A3B in the Q6_K GGUF format, with the shared BF16 vision projector for multimodal prompts.
Links
Tags
Repository: localaiLicense: mit
Cyber-Tiel-Coder is a 35B mixture-of-experts coding model with 3B active parameters, based on Huihui's abliterated Ornith-1.5. This UD-Q4_K_XL build includes MTP speculative decoding, the embedded Sharp chat template, and a BF16 vision projector.
Links
Tags
Repository: localaiLicense: mit
Cyber-Tiel-Coder is a 35B mixture-of-experts coding model with 3B active parameters, based on Huihui's abliterated Ornith-1.5. This UD-Q8_K_XL build includes MTP speculative decoding, the embedded Sharp chat template, and a BF16 vision projector.
Links
Tags
Tiel-Coder-35B-A3B is a 35B-parameter mixture-of-experts model for coding, reasoning, tool use, and vision tasks. This default entry uses the Q4_K_XL GGUF and BF16 vision projector.
Links
Tags
Tiel-Coder-35B-A3B in Q4_K_XL format with MTP speculative decoding and a BF16 vision projector.
Links
Tags
Tiel-Coder-35B-A3B in Q5_K_XL format with MTP speculative decoding and a BF16 vision projector.
Links
Tags
Tiel-Coder-35B-A3B in Q6_K_XL format with MTP speculative decoding and a BF16 vision projector.
Links
Tags
Tiel-Coder-35B-A3B in Q8_K_XL format with MTP speculative decoding and a BF16 vision projector.
Links
Tags
Tiel-Coder-35B-A3B in the higher-quality Q8_K_XL GGUF format, with the BF16 vision projector for multimodal prompts.
Links
Tags
Repository: localaiLicense: openmdw-1.1
NVIDIA Nemotron 3.5 Lightning is a text-only hybrid Mamba-2, attention, and mixture-of-experts model with 30B total parameters and 3B active parameters. It targets reasoning, coding, tool use, multilingual chat, and long-context agent workflows, with a context window of up to one million tokens. This entry uses the official Q4_K_M GGUF. Automatic variant selection can choose the smaller NVFP4 build or the higher-quality Q8_0 build when it fits.
Links
Tags
NVIDIA Nemotron 3.5 Lightning 30B-A3B in the official NVFP4 GGUF format. This is the smallest linked build and retains the model's reasoning, coding, tool-use, multilingual, and long-context capabilities.
Links
Tags
NVIDIA Nemotron 3.5 Lightning 30B-A3B in the official high-quality Q8_0 GGUF format for hosts with enough memory.
Links
Tags
Repository: localaiLicense: apache-2.0

Empero's Qwen3.8 distillation into Qwen3.6-35B-A3B has 35B total parameters and about 3B active per token. This Q4_K_M GGUF build uses the embedded reasoning template and includes the F16 vision projector. Vision is inherited from the base and was not evaluated by the publisher.
Links
Tags
Repository: localaiLicense: apache-2.0

Empero's Qwen3.8 distillation into Qwen3.6-35B-A3B has 35B total parameters and about 3B active per token. This Q5_K_M GGUF build uses the embedded reasoning template and includes the F16 vision projector. Vision is inherited from the base and was not evaluated by the publisher.
Links
Tags