Ling-2.6-flash (104B MoE)
Catalogue summary: InclusionAI's MIT-licensed instruct MoE optimized for fast agent workloads. 104B total parameters, only 7.4B active, hybrid linear attention, 262K context and strong tool-use / multi-step execution with high token efficiency.
Repository editorial metadata; verify comparative claims in the linked upstream material.
Only verified options can be opened. Unsupported apps are clearly marked.
Verified source status
LocalClaw verified a public model card for Ling-2.6-flash (104B MoE) on 2026-09-07, but no public GGUF file in that repository. The catalogue RAM and quantization fields are estimates, not a verified install path.
Only the public model card was verified; no public GGUF file was verified in that repository. Open the model card to confirm current artefacts and supported runtimes. LocalClaw does not claim a one-click LM Studio install.
Source availability
Catalogue record
- Family: ling
- Parameters: 104B (7.4B active)
- Recommended quantization: Q4_K_M
- Catalogue minimum RAM: 80 GB
- Catalogue model size: 65 GB
- Tags: chat, code, reasoning, speed, quality
Practical limits
- Catalogue RAM is a minimum estimate, not a guarantee for every context length or runtime.
- Speed and memory use vary by quantization, backend, context length and system headroom.
- Verify architecture, licence and usage restrictions in the linked upstream material before deployment.
Catalogue tags
- chat
- code
- reasoning
- speed
- quality
Capability profile
Repository catalogue ratings used by LocalClaw's editorial rubric. They are not a standardized third-party benchmark.
Technical notes
Related catalogue entries
Linked mechanically by family, shared tags and nearby RAM tier; this is not a quality ranking.