[P2] Add 50B footprint, complete evaluation table, and multimodal section Critical additions: - Change parameter footprint to 50B release label (was missing) - Add COMPLETE evaluation table comparing Tall vs Qwen3.6-35B-A3B (7 benchmarks) - Add Multimodal Behavior section with 5 benchmark results Core updates: - Add arXiv:2608.09819 tag and paper link - Update citation from blog to arXiv paper - Add contact email New sections (T3-T8): - Routing Behavior and Cost (Tall-specific latency: 1.76s vs Venti 4.68s) - Limitations (base-vs-system comparison, multimodal inherited not tuned) - Safety (no standalone eval, data governance, deployment guidance) - Hardware Requirements (local deployment focus, ~50B footprint) - Training Details (rank 64, alpha 128, expert parameters) - Parameter count clarification (50B label vs 35B base)

#5
Ready to merge
This branch is ready to get merged automatically.

Sign up or log in to comment