Maple-Preview vPreview
Open Source
reasoning
ternary
Maple-Preview is an open-source 20B-A1B ternary-weight reasoning large language model optimized for efficient on-device inference. It features a 24-layer, 256-expert architecture designed to deliver state-of-the-art reasoning capabilities while maintaining high performance on consumer hardware like the Mac mini M4.
20B-A1B ternary-weight architecture
256-expert configuration with 8 active experts
131,072 token context window
PricingFree (Open Source)
VersionPreview