flash-lite is more of their luna tier competitor but even still not quite there yet, but gemini's dominance on multimodal and image understanding i think really gets downplayed on this site when most people think the only think you can do with LLMs is write code
Flash models and Gemini make more sense when you consider Gemini Enterprise and Workspace. Oh HN we generally care a lot about writing software. However, until Fable, Gemini 3.1 Pro was my default for doing any sort of discussion outside of software engineering. Fable is now on par with things like modifying cars, etc. But I am guessing Fable is a /lot/ more expensive to use.
And in a typical enterprise environment dealing with documents, images, and broader business reasoning skills matter. Agents are not just for code and text :)
That's been my association as well. I see Flash get brought up a lot in relation to things like OCR and PDF processing frequently, and a lot of other routine multimodal workloads.
If anyone knows of a cheaper vision llm with the same accuracy I would love to switch
Yes this is my impression as well. To be fair I didn't compare to Luna yet, but Gemini 3.5 Lite is a very good and cheap multi-modal data extraction model.
Ultimately it would track that in the real world, people will want to point cameras at things and get answers.
I pay for ChatGPT and Gemini, and while Sol is a total beast with anything text, it still poisoned my cucumber bed. Which I will be bitter about for at least a few years while the bed recovers. Gemini (even flash) is exceptionally talented at viewing photos and telling you what to do/what it is (and telling me I just misidentified the problem with my cucumbers and spraying off the "bugs" actually just spread the bacteria everywhere.)