Apps & Tools
Ling-3.0-tiny
A fast 7.9B reasoning MoE model optimized for local inference on Macs and GPUs.
byAnt Ling
Playable
Ling-3.0-tiny media is blocked
Allow external media to connect to the provider and play this content.
Description
Ling-3.0-tiny is a lightweight hybrid reasoning Mixture-of-Experts (MoE) model with 7.9B total parameters and 1.3B activated parameters per token. It features a native 'thinking mode' for multi-step reasoning and is optimized for local deployment, achieving up to 90 tokens/s on Apple Silicon (M4 Pro) and 160 tokens/s on high-end GPUs.
Descriptions, tags, and model credits may be AI-generated or inferred from public sources and can be incomplete or wrong. Learn more.