2 days ago • 4 min read Qwen 2.5 Max Architecture and Benchmarks: Alibaba's Frontier Open Weights Alibaba's Qwen 2.5 series sets record-breaking scores on MATH, HumanEval, and multilingual comprehension, challenging Silicon Valley hegemony. By Marcus Vance 11,201 views
4 days ago • 5 min read The Rise of Reasoning Models: How Test-Time Compute (o1, o3, R1) Replaces Pure Scaling Why pre-training scaling laws hit a wall of diminishing returns, and how test-time compute search graphs unlocked new... By Marcus Vance 16,101 views
6 days ago • 4 min read Prompt Engineering in 2026: Cognitive Scaffolding, DSPy, and System Prompts Why manual trial-and-error prompting is being replaced by compiled program graphs like DSPy, structured output schemas, and automated... By Marcus Vance 11,901 views
1 week ago • 5 min read Transformer Alternatives: Mamba, State Space Models (SSMs), and Hybrid Attention Can linear-time State Space Models overcome the quadratic memory bottleneck of standard self-attention? Architectural comparison of Mamba-2 and... By Marcus Vance 9,901 views