Same prompt presented to two Google models, the Gemini 3.1 Pro generally used for tasks that require reasoning and judgment and the less expensive, lighter, speedier Gemini 3 Flash. The pricier Pro model finished in 85 steps, whereas the less expensive Flash model went through nearly 1,000 steps…then failed. Cost: The less expensive model failed after running up $14 worth of token use, while the pricier one succeeded for just $1. Even the same model might use a different number of tokens each time it completes a task. Anyone observed this ? submitted by /u/XIFAQ
Originally posted by u/XIFAQ on r/ArtificialInteligence
You must log in or # to comment.
