I'm running lexical analysis on gemini 3.8 flash and the latency progress is incredible. For my tasks the latency is reduced by ~40% w/ quality on par.
experimenting w/ flash 3.7 on morphology and it looks promising. google models looks capable on linguistic side and prose. are there any reliable benchmarks on this?
reply