Google launches Nano Banana 2 Lite image model and expands Gemini Omni Flash to developers

Google has announced Nano Banana 2 Lite, its fastest and most cost-efficient Gemini Image model to date, alongside expanded developer access to Gemini Omni Flash, its multimodal video generation and conversational editing model. Continue reading “Google launches Nano Banana 2 Lite image model and expands Gemini Omni Flash to developers”

Google unveils Gemma 4 12B for local AI agents, coding, and multimodal reasoning

Google DeepMind has introduced Gemma 4 12B, a new open-weight multimodal model designed to bring agentic intelligence directly to laptops with mobile-first efficiency and advanced reasoning. Continue reading “Google unveils Gemma 4 12B for local AI agents, coding, and multimodal reasoning”

Google expands Deep Research and Deep Research Max for autonomous research workflows

Google has introduced Deep Research and Deep Research Max, powered by Gemini 3.1 Pro, marking a step change in its autonomous research agents. The update builds on the Gemini Deep Research preview released in December via the Interactions API, expanding it into a system capable of generating structured, fully cited reports. Continue reading “Google expands Deep Research and Deep Research Max for autonomous research workflows”

OpenAI rolls out ChatGPT Images 2.0 with improved text rendering, multilingual support, and thinking capabilities

OpenAI has introduced ChatGPT Images 2.0, a next-generation image generation model designed to produce precise, structured, and usable visual outputs. The model is built to handle complex visual tasks, improve instruction accuracy, and generate images that better reflect real-world design needs. Continue reading “OpenAI rolls out ChatGPT Images 2.0 with improved text rendering, multilingual support, and thinking capabilities”

Meta introduces Muse Spark AI model with multimodal reasoning and multi-agent capabilities

Meta has introduced Muse Spark, a new AI model developed by Meta Superintelligence Labs. The model is part of a broader Muse series designed to advance multimodal understanding, reasoning, and agent-based workflows, as part of Meta’s effort toward personal superintelligence. Continue reading “Meta introduces Muse Spark AI model with multimodal reasoning and multi-agent capabilities”

Microsoft rolls out MAI-Transcribe-1, MAI-Voice-1 and MAI-Image-2 in Foundry public preview

Microsoft has announced three new AI models—MAI-Transcribe-1, MAI-Voice-1, and MAI-Image-2—in public preview through its AI development platform Microsoft Foundry. Continue reading “Microsoft rolls out MAI-Transcribe-1, MAI-Voice-1 and MAI-Image-2 in Foundry public preview”

Google unveils Gemma 4 open AI models with advanced reasoning and agentic workflows

Google has introduced Gemma 4, an open model family designed for advanced reasoning and agentic workflows. The models are built to deliver high intelligence-per-parameter while remaining efficient across a wide range of hardware environments. Gemma 4 is positioned as part of Google’s broader AI ecosystem and complements its proprietary Gemini models. Continue reading “Google unveils Gemma 4 open AI models with advanced reasoning and agentic workflows”

Google opens applications for 2026 Startup Accelerator India cohort

Google has opened applications for the 2026 cohort of its startup support initiative, the Google for Startups Accelerator: India. The program is designed for AI-first startups in India that are developing scalable solutions across advanced AI domains and moving toward real-world deployment. Continue reading “Google opens applications for 2026 Startup Accelerator India cohort”

Google rolls out Gemini 3.1 Flash Live for real-time voice AI conversations, expands Search Live globally

Google has introduced Gemini 3.1 Flash Live, a real-time audio and voice AI model designed to enable faster, more natural conversational experiences. The model enhances latency, reliability, and dialogue quality for developers, enterprises, and everyday users, supporting the next generation of voice-first and multimodal AI applications. Continue reading “Google rolls out Gemini 3.1 Flash Live for real-time voice AI conversations, expands Search Live globally”

Google rolls out Gemini 3 with Deep Think mode, enhanced coding and agentic actions

Google has introduced Gemini 3, the next generation of its AI model series, bringing upgrades in long-form reasoning, multimodal interpretation, interface generation, developer tools, and agent-based task execution. Continue reading “Google rolls out Gemini 3 with Deep Think mode, enhanced coding and agentic actions”