Basically the title itself, also i doesn’t even answer in the last, it just throws an error. I get that newer models often use internal reasoning/planning, but this feels excessive for such a straightforward request. A Flash model is supposed to prioritize speed and responsiveness, so seeing it spend noticeable time “thinking” before answering a basic question is surprising. Benchmarks are making lower and lower sense now, I like to think that even opus 4.6 smarter in terms of raw intelligence. these newer models may generate better quality code and may have better information but usually they overdo on some basic stuff. Prompting any gemini model through web is a nightmare. even on basic ppt generation stuff. I saw that the newer flash models performed worse than 3.1 pro. Can anyone tell me what the reason might be? How and which one to use without having to deal with this bs again. submitted by /u/nahmanhuh
Originally posted by u/nahmanhuh on r/ArtificialInteligence
