THE METER
SECTION 03
ISSUE 001
Mixture-of-Experts Goes Application-Native
Projection: Meta demonstrates mixture-of-experts routing inside a native multimodal model, while Qwen 3.7 combines Max and Plus tiers with selectable thinking modes. The next routing boundary may sit in the application, choosing modules, tools, and policies. It matters only if workload evaluations beat a simpler fixed route.
Why this idea is here
What the evidence establishes.
Meta documents native multimodality, mixture-of-experts routing, teacher distillation, and a claimed ten-million-token context model; Qwen 3.7 documents hybrid-thinking controls across current Max and Plus tiers. These are source-backed premises for this inference; they do not by themselves prove broad adoption or the eventual outcome.
Source ledger
Read the sources.
- S01Current Llama models and resources
official live model documentation / dated Live source · verified 2026-07-10 / retrieved 2026-07-10
- S02Qwen 3.7 deep-thinking and hybrid-thinking models
official current model documentation / published 2026-06-18 / retrieved 2026-07-10