Google has released its latest artificial intelligence image generation tools, making the new Gemini API models generally available for developers worldwide. Specifically, the release includes Nano Banana Pro and Nano Banana 2 to support production-ready workflows. These tools provide stable options for creators who require precise control or high-volume output.

Specifications of the New Gemini API Models

The two options target different apps development needs. For instance, Nano Banana 2, also known as Gemini 3.1 Flash Image, focuses on speed and cost-effectiveness. Consequently, developers can use this option for high-volume workflows that require rapid processing. This model is designed to handle large-scale tasks without incurring high operational costs.

Meanwhile, Nano Banana Pro, or Gemini 3 Pro Image, serves complex creative workflows. This model provides professional-grade precision for detailed design tasks. As a result, users can achieve higher accuracy in their visual outputs. It is particularly suited for projects that demand exact details and high-quality rendering.

Core Capabilities and Features

Both Gemini API models share several core technical strengths. Notably, they maintain subject consistency across five distinct characters in an image. This feature allows developers to create sequential visual content while keeping character identities stable. In addition, the models support accurate text rendering and localization, which helps in generating readable text within images across different languages.

Furthermore, developers can utilize 4K upscaling and custom aspect ratios to fit various display formats. The system accepts up to 14 reference image inputs for better guidance during the generation process. Crucially, grounding with Google Search ensures factual accuracy in the generated images, reducing errors in visual representations. This grounding capability helps verify real-world details before rendering the final output.

Current Developer Integrations

Several smart-devices platforms have already integrated these Gemini API models into their systems. For example, builders use the technology for 3D mesh generation with SplineTool. Meanwhile, other developers use Krea_ai to enable real-time editing features. These integrations demonstrate how the models function in active development environments.

By utilizing these platforms, creators can test the limits of the new image generation capabilities. The combination of 3D tools and real-time editing shows the versatility of the API. Consequently, more applications are expected to adopt these models in the coming months.

Production Readiness and Access

The models are now stable and ready for production environments. Developers can access them directly through the official API. This release allows businesses to scale their image generation tasks with predictable performance. In conclusion, the availability of these tools provides developers with reliable options for building visual applications. Interested parties can explore the official documentation to begin integration.

Source: X (@googledevs)