NewsHK

Google's Gemini 3.8 Flash arrives as AI token pricing competition sharpens across the sector

9/3/2026

The AI inference market is repricing noticeably, and the cadence of model releases has accelerated to match.

Google has released Gemini 3.8 Flash, its third Flash-series model in six weeks, and the company is pitching introductory API rates of $0.75 per million input tokens and $3.75 per million output tokens through the end of the year.

Two variants ship simultaneously.

The standard Gemini 3.8 Flash is positioned as a general-purpose workhorse suited to agentic tasks and software development, and Google describes the combined package as its best reasoning and coding model to date.

Keep reading

Read the full story

Open on NewsHK