Kimi K3 basically matches US frontier labs, Deepseek V4 is ~90% of frontier intelligence at ~5% of the cost, yet MAI team (one of the most well-resourced AI teams in the world) won’t submit MAI-Thinking-1 or MAI-Code-Flash to Artificial Analysis for benchmarking, which is a telling sign of how far behind they are. I understand that MAI was first focused on lowering COGS for MS teams transcripts / image generation for Copilot (their audio and image models are at the frontier and super cost-effective, see them on Artificial Analysis), but being this far behind on coding and general intelligence is quite pathetic given their resources. Not sure what Satya is thinking. MSFT stock is likely stuck until they can put out a model that benchmarks well submitted by /u/NormandyPark0
Originally posted by u/NormandyPark0 on r/ArtificialInteligence
