xAI has launched Grok Imagine Image 2.0 with region-level editing, multi-image references and improved creative control as competition with OpenAI intensifies.
Alibaba has launched Qwen3.8-Max, its most advanced AI model, claiming breakthrough performance as competition with leading US AI companies intensifies.
ByteDance has launched Seedance 2.5, an AI video model supporting 50 multimodal inputs, 30-second video generation and advanced editing capabilities.
Google has launched Gemini 3.6 Flash and Gemini 3.5 Flash-Lite, introducing faster, more efficient AI models designed for developers, enterprise applications and AI agents.
Mira Murati's Thinking Machines has launched Inkling, an open weight multimodal AI model designed for enterprise customisation and developer flexibility.
AI video platform PixVerse has raised $439 million in a Series A extension to expand global operations and strengthen its generative AI video technology.
Google has selected 20 Indian AI startups for its 2026 accelerator programme, offering mentorship, Gemini AI access and Google Cloud support to help them scale globally.
Google has launched Nano Banana 2 Lite and Gemini Omni Flash, introducing faster AI image generation and multimodal video creation tools for creators and enterprises.
IndiaAI-backed AvataarAI has launched Varya, an indigenous video generation model designed to create high-quality AI videos from text and image inputs.
Google has introduced Gemini Omni, a multimodal AI model focused on advanced video creation, editing, and generative media workflows.
OpenAI has launched new voice intelligence features in its API to support advanced conversational and multimodal AI applications.
Apple reportedly faces a $250 million AI marketing settlement while planning multi-model AI options for future iOS updates.
ChatGPT Images 2.0 could reshape visual marketing by enabling faster, scalable, and localised creative development across campaigns and platforms.
LG introduces EXAONE 4.5, a multimodal AI model designed to power next-generation intelligent systems across industries.
OpenAI is reportedly planning to integrate its Sora AI video generation model into ChatGPT, expanding the platform’s capabilities for multimedia content creation.
Perplexity has launched voice mode for Perplexity Computer, enabling users to interact with its AI platform through spoken commands and hands free computing.
Google rolls out Nano Banana 2 as its default AI image generation tool, enhancing prompt accuracy, visual quality and enterprise integration capabilities.
OpenAI is reportedly developing a new voice model as it prepares for an AI hardware launch, signalling a stronger push into multimodal and voice-based AI systems.
Aionos expands multimodal AI and agent based solutions as its CTO forecasts rapid adoption across healthcare and enterprise sectors.
Healthify launches Ria Voice, a real time multimodal AI health coach built with OpenAI to offer personalised voice based wellness guidance.
Google unveils Gemini 3, stating major improvements in reasoning, math, coding and multimodal tasks, positioning it ahead of leading AI models in global benchmark tests.
Multimodal AI is reshaping customer engagement by combining text, voice, image, and video insights to deliver real-time, personalized, and emotionally intelligent brand interactions across industries.
Agora partners with OpenAI to integrate Realtime API, enabling multimodal AI agents that deliver seamless real-time voice, video, and text interactions for enterprises.