https://preview.redd.it/hthhgu6b74fh1.jpg?width=1024&format=pjpg&auto=webp&s=9379c13f0f0671760c5565c560f369ca861996f5 Two mega-models dropped back-to-back this July: Moonshot AI’s Kimi K3 (2.8T open-weight MoE) and Alibaba Cloud’s Qwen 3.8 Max (2.4T sparse MoE) . Both are redefining what “frontier AI” means. Kimi K3 → Open-weight, 2.8T parameters, “always-on” reasoning, 90% prompt caching. Perfect for self-hosted enterprise setups and rapid synchronous dev workflows. Qwen 3.8 Max → Multimodal (text, image, video, PDF), async test-time compute loops (30–80 min), protocol-fluid APIs. Acts more like an autonomous worker than a chatbot. Verdicts from real-world scenarios: Codebase refactoring → Kimi K3 wins (speed + caching efficiency). One-shot full-stack app dev → Qwen 3.8 Max wins (autonomous Playwright validation). Financial chart + video ingestion → Qwen 3.8 Max wins (native multimodal). TL;DR: Kimi K3 = speed + cost control. Qwen 3.8 Max = autonomy + multimodality. submitted by /u/Remarkable-Dark2840
Originally posted by u/Remarkable-Dark2840 on r/ArtificialInteligence
