Google

2026Google

Gemini 3.5 Live Translate v3.5

Paid
AI
Translation

Gemini 3.5 Live Translate is Google's latest audio model that delivers near real-time, natural-sounding speech-to-speech translation in over 70 languages. It automatically detects languages and generates smooth translations that preserve the speaker's intonation, pacing, and pitch, enabling fluid communication across language barriers.

Near real-time speech-to-speech translation in over 70 languages
Automatic language detection
Preserves speaker's intonation, pacing, and pitch
Version3.5
RegionUnited States
2026Google

LiteRT.js

Free
Web AI
Machine Learning

LiteRT.js is a high-performance web AI runtime that enables developers to run machine learning models directly in the browser. By leveraging WebAssembly and hardware acceleration like WebGPU and WebNN, it allows for low-latency, private, and serverless AI inference using .tflite models.

Native hardware acceleration via XNNPACK (CPU), WebGPU (GPU), and WebNN (NPU)
Seamless deployment of existing .tflite models in web applications
Tailored quantization support via AI Edge Quantizer for optimized performance and model size
PricingFree & Open-source
RegionUnited States
2026Google

DiffusionGemma v1.0

Open Source
Featured
AI
Open Source

DiffusionGemma is an experimental open model developed by Google that utilizes text diffusion techniques to generate entire blocks of text simultaneously, achieving up to four times faster text generation on dedicated GPUs compared to traditional autoregressive models. Released under an Apache 2.0 license, this 26-billion parameter Mixture of Experts (MoE) model is designed for researchers and developers working on speed-critical, interactive local workflows such as in-line editing, rapid iteration, and generating non-linear text structures.

Up to 4x faster text generation on dedicated GPUs
Generates entire blocks of text simultaneously
Bi-directional attention allowing every token to attend to all others
PricingFree
Version1.0
2026Google

TranslateGemma v1.0

Open Source
AI
Machine Translation

TranslateGemma is a suite of open machine translation models developed by Google AI, built upon the Gemma 3 architecture and designed to support 55 languages. The models are available in 4B, 12B, and 27B parameter sizes, optimized for deployment across various devices, from mobile and edge hardware to laptops and cloud instances.

Open-source machine translation models
Supports 55 languages
Available in 4B, 12B, and 27B parameter sizes
PricingFree
Version1.0
2025Google

Conductor v1.0.0

Open Source
Coding

Conductor is an extension for Gemini CLI that facilitates context-driven development by enabling developers to create formal specifications and plans alongside their code in persistent Markdown files. This approach allows for planning before building, reviewing plans prior to coding, and maintaining developer control throughout the development process.

Plan before you build: Create specifications and plans that guide the agent for new and existing codebases.
Maintain context: Ensure AI follows style guides, tech stack choices, and product goals.
Iterate safely: Review plans before code is written, keeping developers firmly in the loop.
PricingFree
Version1.0.0
2026Google

Gemini CLI v1.6 Preview vv1.6 Preview

Open Source
Coding
Robotics

Google updated Gemini CLI to v1.6 Preview with improved spatial reasoning capabilities, specifically targeting developers who use AI for hardware and robotics-related coding. The terminal-first agent mode now handles spatial concepts more accurately, making it a go-to tool for embedded systems and robotics software engineers.

Improved spatial reasoning for hardware/robotics coding
Terminal-first AI agent mode
Enhanced accuracy for embedded systems development
PricingFree with Google AI account
Versionv1.6 Preview
2026Google

Gemini 3.7 Flash v3.7 Flash

Paid
Featured
AI Models
Coding

Gemini 3.7 Flash is an advanced, high-performance AI model optimized for coding, agentic workflows, and complex reasoning tasks. It delivers significant improvements in debugging, issue resolution, and web development while maintaining cost-efficiency for developers and enterprises.

Enhanced coding capabilities for debugging and production-ready code generation
Improved reasoning and accuracy for knowledge-dense fields like law and finance
Optimized for agentic workflows with better multi-step planning and tool calling
Version3.7 Flash
RegionUnited States
2026Google

Google AI Studio Android App Builder vLaunch (May 2026)

Paid
Google
Android

Google AI Studio can now build entire native Android apps from a single prompt — no installation required. Powered by Gemini, it generates Kotlin/Jetpack Compose apps with an embedded browser-based Android emulator for live preview, direct device install via USB/ADB, and one-click publish to Google Play internal testing track.

Prompt-to-Android-app in the browser — generates full Kotlin/Jetpack Compose apps with an embedded Android emulator for live preview; no SDK, no local setup, supports Camera, GPS, Bluetooth, NFC, and Gemini API integrations out of the box
Instant device install and Google Play publishing — connect your Android phone via USB and install directly from AI Studio using ADB; publish to a Google Play internal testing track in minutes with automatic bundle packaging and app record creation
Seamless handoff to Android Studio or Antigravity — download a ZIP or export directly to GitHub to continue development locally; integrates with Firebase (Firestore, Auth, App Check) coming soon and supports Gemini in Android Studio for advanced workflows
PricingFree (Google AI Studio)
VersionLaunch (May 2026)
2026Google

Google Colab CLI v1.0

Open Source
Featured
Google Colab
CLI

The Google Colab Command-Line Interface (CLI) is a tool that bridges local terminals and remote Colab runtimes, enabling developers and AI agents to execute scripts, download models, and automate machine learning pipelines seamlessly. It offers features like instant GPU/TPU provisioning, remote execution of Python scripts, artifact recovery, and interactive access to remote environments.

Zero-Friction Accelerator Provisioning: Request high-powered GPUs or TPUs instantly (e.g., `colab --gpu A100` or `colab --gpu T4`).
Simple Remote Execution: Run local Python scripts and complex ML pipelines directly on Colab runtimes using `colab exec`.
Seamless Artifact Recovery: Retrieve models, datasets, and replayable `.ipynb` logs via `colab download` and `colab log`.
PricingFree
Version1.0
2026Google

Angular v22.1.1

Open Source
web-framework
typescript

Angular is a comprehensive, open-source development platform and framework for building scalable, high-performance web applications. It leverages TypeScript to provide a robust architecture for mobile and desktop web development, maintained by a dedicated team at Google.

Component-based architecture for modular code
Angular CLI for efficient project scaffolding and management
Built-in support for Server-Side Rendering (SSR) and hydration
PricingFree and open-source (MIT License)
Version22.1.1
2026Google

Antigravity CLI v2.0 (May 2026)

Paid
Featured
Google
CLI

Antigravity CLI is Google's terminal-first agentic coding tool — the successor to Gemini CLI — powered by Gemini 3.5 Flash. Launched at Google I/O 2026 as part of Antigravity 2.0, it lets developers orchestrate multi-agent workflows, schedule background tasks, and build custom agents from the terminal with full Google Cloud, Android, Firebase, and AI Studio integration.

Terminal-first multi-agent orchestration — run parallel subagent workflows, schedule background tasks, and design custom agent pipelines from the CLI; powered by Gemini 3.5 Flash and successor to Gemini CLI with full migration path provided
Native voice commands and Antigravity SDK — add voice-driven coding via terminal, build custom agents with the new Antigravity SDK, connect Google Cloud projects, and use custom agent templates in AI Studio for enterprise workflows
Deep integration with Google ecosystem — export projects from AI Studio to Antigravity, integrate with Android CLI commands, Firebase, and Google Cloud; available on AI Ultra plan ($100/mo for 5x limits, $200/mo for 20x limits)
Version2.0 (May 2026)
RegionGlobal
2026Google

Nano Banana 2 Lite v2 Lite

Freemium
AI
Image Generation

Nano Banana 2 Lite is Google's most efficient Gemini Image model, designed for rapid image generation and editing with minimal latency and cost. It offers high-quality outputs while maintaining the control and accuracy expected from the Nano Banana series.

Lightning-fast latency
Cost-efficient at scale
No compromise on quality
Pricing$0.034+ per 1k resolution image
Version2 Lite
2026Google

Gemini CLI 0.33.0 v0.33.0

Open Source
Agentic AI
Coding

Gemini CLI is an open-source command-line interface developed by Google that integrates the power of Gemini AI models directly into your terminal. It enables developers to interact with Gemini models seamlessly, enhancing productivity and workflow efficiency. v0.33.0 introduces Plan Mode: Introducing a read-only environment for researching, designing and planning before implementing. A read-only mode that allows you to safely explore and work with Gemini CLI to come up with a plan prior to implementation. Plan mode can leverage read-only MCP tools, Agent Skills and is built to be fully extensible

Integration with Gemini AI models
Command-line interface for AI interactions
Open-source and community-driven development
PricingFree
Version0.33.0
2026Google

Magika v1.1.0

Open Source
AI
File Detection

Magika is an AI-powered file type detection tool that utilizes deep learning to provide fast and accurate identification of file contents. It is designed to process hundreds of billions of files weekly, supporting both binary and textual formats with approximately 99% accuracy.

AI-powered detection with ~99% accuracy across 200+ content types
High-performance inference (~5ms per file) on a single CPU
Support for command line, Python, JavaScript/TypeScript, and Go
PricingFree & Open-Source
Version1.1.0
2026Google

Google Code Wiki v2026-05-08

Free
Gemini
AI Documentation

Code Wiki is a new perspective on development for the agentic era. Gemini-generated documentation that's always up-to-date. Automatically generates and maintains interactive knowledge bases from code repositories, with AI agents that update docs when PRs merge, diagrams that visualize architecture, and natural language chat to understand your codebase.

Gemini-generated documentation: AI agent automatically generates and maintains a rich, interactive knowledge base from your code
Always up-to-date: Every time a pull request is merged, the relevant documentation is automatically updated
Architecture diagrams: Complex systems transformed into clear, intuitive visuals that bring architecture to life
PricingFree
Version2026-05-08
2026Google

WaxalNLP v1.0.0

Open Source
Audio

The WaxalNLP dataset is a comprehensive collection of audio data for both Automated Speech Recognition (ASR) and Text-to-Speech (TTS) tasks in 14 African languages. It aims to enhance the accuracy and fluency of speech and language technologies for underserved African languages and serves as a resource for digital preservation.

Contains approximately 1,250 hours of transcribed natural speech for ASR
Includes about 240 hours of scripted natural speech for TTS
Covers 14 African languages spoken by over 100 million people across 40 Sub-Saharan countries
PricingFree
Version1.0.0
2026Google

Agentic Vision in Gemini 3 Flash vGemini 3 Flash

Open Source
CV
Agentic AI

Agentic Vision is a new capability in Google's Gemini 3 Flash model that combines visual reasoning with code execution to ground answers in visual evidence. This feature enables the model to actively manipulate and analyze images, enhancing its ability to provide accurate and contextually relevant responses.

Visual reasoning combined with code execution
Active image manipulation and analysis
Enhanced accuracy in complex image-related queries
VersionGemini 3 Flash
RegionUnited States
2026Google

ADK for Go 2.0 v2.0

Open Source
AI Agents
Go

ADK for Go 2.0 is a framework for building complex, multi-agent Go applications using a graph-based workflow engine. It provides built-in support for human-in-the-loop interaction, dynamic orchestration, and state management within an idiomatic Go environment.

Graph-based workflow engine for multi-agent orchestration
Built-in human-in-the-loop (HITL) support with durable state resumption
Dynamic orchestration capabilities using standard Go code
PricingFree
Version2.0
2026Google

Google Antigravity SDK v0.1.6

Open Source
AI Agents
Python SDK

The Google Antigravity SDK is a Python library that provides developers with programmatic access to the Google Antigravity agent harness. It allows for the creation, testing, and deployment of autonomous AI agents using the same core infrastructure and tools that power Google's Antigravity 2.0 and CLI products.

Unified agent harness used by Antigravity 2.0 and CLI
Support for custom Python tools and MCP server integration
Declarative safety policy engine for granular control
PricingFree
Version0.1.6
2026Google

Google Workspace CLI

Open Source
AI Agents
Agentic AI

The Google Workspace CLI is a command-line tool developed by Google that enables users to manage various Google Workspace services, including Drive, Gmail, Calendar, Sheets, Docs, Chat, and Admin, directly from the terminal. It is dynamically built from Google's Discovery Service and incorporates AI agent capabilities to enhance user experience.

Manage Google Workspace services via command-line interface
Dynamically built from Google's Discovery Service
Incorporates AI agent capabilities for enhanced user experience
PricingFree
RegionUnited States
2026Google

Universal Commerce Protocol (UCP) v1.0

Open Source
AI
Open Source

UCP is an open-source standard developed by Google to facilitate seamless agentic commerce experiences. It enables AI agents and merchant systems to communicate effectively, allowing users to complete purchases within a single conversation. By providing a shared language, UCP eliminates the need for custom integrations between retailers and various platforms. ([marktechpost.com](https://www.marktechpost.com/2026/01/12/google-ai-releases-universal-commerce-protocol-ucp-an-open-source-standard-designed-to-power-the-next-generation-of-agentic-commerce/?utm_source=openai))

Standardizes communication between AI agents and merchants
Supports various commerce verticals like shopping, travel, and services
Integrates with payment and credential providers
PricingFree
Version1.0
2025Google

A2UI v0.8

Open Source
AI Agents
Agentic AI

A2UI is an open-source specification and set of libraries developed by Google that enables agents to describe rich native interfaces in a declarative JSON format. This approach allows client applications to render these interfaces using their own components, facilitating secure and interactive user interfaces across trust boundaries without the need to send executable code.

Declarative JSON format for UI descriptions
Security-focused design to prevent arbitrary code execution
Framework-agnostic, supporting various client applications
PricingFree
Version0.8
2026Google

Gemini 3.1 Pro v3.1 Pro

Paid
AI Agents
Agentic AI

Gemini 3.1 Pro is Google's latest AI model designed to tackle complex tasks requiring advanced reasoning. It offers improved performance over its predecessor, Gemini 3 Pro, with a verified score of 77.1% on the ARC-AGI-2 benchmark, more than doubling the reasoning performance of Gemini 3 Pro. ([blog.google](https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-1-pro/?utm_source=openai))

1 million token context window
Advanced reasoning capabilities
Integration with Google Antigravity for agentic development
Version3.1 Pro
RegionUnited States
2026Google

Gemini 3.6 Flash v3.6 Flash

Paid
AI Models
Developer Tools

Gemini 3.6 Flash is a high-performance, cost-efficient AI model optimized for coding, knowledge work, and complex multimodal tasks. It builds upon the capabilities of its predecessor, 3.5 Flash, by reducing token usage by up to 17% and streamlining multi-step reasoning workflows.

Up to 17% reduction in output token usage compared to Gemini 3.5 Flash
Enhanced coding precision with fewer execution loops and unwanted edits
Streamlined multi-step reasoning and tool call execution
Version3.6 Flash
RegionUnited States
2026Google

WebMCP vEarly Preview

Open Source
ML
MCP

WebMCP is an initiative by Google to standardize how AI agents interact with websites, aiming to enhance the efficiency and reliability of these interactions. By defining structured tools, WebMCP enables AI agents to perform actions on websites with increased speed and precision.

Declarative API for standard actions
Imperative API for complex interactions
Enhanced AI agent workflows
PricingFree
VersionEarly Preview
2026Google

Google Antigravity 2.0 v2.0

Freemium
AI Agents
IDE

Google Antigravity 2.0 is an agentic development platform designed to orchestrate multiple autonomous AI agents across complex software projects. It acts as a central command center for developers, enabling parallel task execution, background automation, and cross-surface integration.

Abstracted UI for centralizing and orchestrating multiple autonomous AI agents
Dynamic Subagents for executing complex, parallelized workflows
Scheduled Tasks for automating routine background operations
Version2.0
RegionUnited States
2026Google

Gemma 4 12B v4 12B

Open Source
Featured
AI
Machine Learning

Gemma 4 12B is a unified, encoder-free multimodal model designed to deliver high-performance AI capabilities directly to laptops. It integrates vision and audio inputs seamlessly into its language model backbone, enabling advanced reasoning and agentic workflows without the need for separate encoders. This model is optimized for local deployment, requiring only 16GB of VRAM or unified memory, making it accessible for developers seeking powerful AI tools on standard hardware.

Unified, encoder-free architecture integrating vision and audio inputs directly into the language model backbone
Advanced reasoning capabilities with performance nearing larger models, enabling multi-step reasoning and agentic workflows
Optimized for local deployment with a memory footprint suitable for laptops with 16GB of VRAM or unified memory
PricingFree
Version4 12B
2026Google

Gemini API Event-Driven Webhooks v2026-05-04

Paid
Gemini API
Google AI

Google has launched event-driven webhook support in the Gemini API, enabling developers to build complex long-running agentic applications without polling. Instead of repeatedly checking job status, developers register a webhook URL and Gemini calls it automatically when long-running tasks complete — reducing latency, cutting infrastructure overhead, and making async AI pipelines significantly simpler to build.

Register a webhook URL with the Gemini API and receive automatic callbacks when long-running jobs complete — eliminates polling loops and reduces wasted compute for agentic pipelines
Purpose-built for complex, long-running agentic applications: model fine-tuning jobs, batch inference, async generation tasks, and multi-step reasoning workflows that take minutes to hours
Drop-in improvement for existing Gemini API integrations — standard HTTP webhook pattern, works with any backend infrastructure, no SDK changes required
PricingIncluded with Gemini API (pay-per-use)
Version2026-05-04
2026Google

Interactions API vGeneral Availability

Free
Featured
AI
API

Google's Interactions API is a unified interface for interacting with Gemini models and agents, enabling developers to build advanced applications with complex interactions, server-side state management, background execution, and multimodal generation.

Unified interface for Gemini models and agents
Server-side state management
Background execution for long-running tasks
PricingFree & Paid Tier
VersionGeneral Availability
2026Google

AX (Agent Executor)

Paid
Agent Runtime
Distributed Systems

AX (Agent Executor) is Google's open-source distributed agent runtime for production AI agents. It coordinates agentic loops across distributed isolated actors (skills, tools, agents), manages durable execution state with automatic recovery and resumption, and provides a single-writer architecture for consistent state management — designed to scale from single machines to Kubernetes clusters.

Distributed agent runtime — controller, skills, tools, and agents execute in isolation as separate actors; event log ensures durable execution state with automatic recovery and resumption even in complex distributed setups; supports branching from any checkpoint via ax fork
Built-in resumption and reliability — single-writer architecture ensures consistent state; clients can reconnect mid-execution and replay missed events; native Kubernetes support via Agent Substrate with higher-density agentic workload scheduling
Model and harness agnostic — supports custom remote agents via gRPC AgentService, MCP tool servers, and any LLM backend; complete auditing and policy control of all user and agentic calls through a central controller; CLI for local or server execution
PricingFree (Open Source, Apache-2.0)
RegionGlobal

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode