Google announced the launch of three new Gemini models designed for improved speed, token efficiency, and reliability at scale. This expansion includes models built to be faster and more reliable, suggesting a focus on optimizing inference workloads.