NewsHK
The AI inference market is repricing noticeably, and the cadence of model releases has accelerated to match.
Google has released Gemini 3.8 Flash, its third Flash-series model in six weeks, and the company is pitching introductory API rates of $0.75 per million input tokens and $3.75 per million output tokens through the end of the year.
Two variants ship simultaneously.
The standard Gemini 3.8 Flash is positioned as a general-purpose workhorse suited to agentic tasks and software development, and Google describes the combined package as its best reasoning and coding model to date.
Keep reading