Topic: Multimodal Capabilities
-
Google NotebookLM Rebrand Could Increase AI Scraping Risks
Google rebranded NotebookLM to Gemini Notebook, replacing the old user agent "Google-NotebookLM" with "Google-GeminiNotebook"; the legacy agent will stop functioning in August 2026, requiring site owners to update hardcoded configurations. Gemini Notebook's Discover Sources feature scrapes online...
Read More » -
Gemini’s new tools give you more control over your phone
Google's "Gemini Intelligence" umbrella, reserved for premium phones like the Galaxy S26 series, bundles new and existing AI features, requiring users to opt in. Task automation expands to more apps and gains multimodality, allowing Gemini to perform actions based on voice, text, or a photo/scree...
Read More » -
Google Launches Gemma 4 AI Model
Google has launched the fourth-generation Gemma AI model family, featuring four distinct sizes for hardware from mobile to workstations and a pivotal shift to the permissive Apache 2.0 license for broader commercial use. The models, which share technology with Google's Gemini 3, include compact e...
Read More » -
ServiceNow and Nvidia Launch Open-Source Security Model
ServiceNow and Nvidia have launched Apriel 2.0, an open-source AI model designed for building custom agents and enhancing security operations in regulated environments. Built on Nvidia’s Nemotron architecture, it integrates with ServiceNow’s compliance-certified platform, allowing seamless use of...
Read More » -
OpenAI's Voice AI Strategy: Expressive Speech for Enterprise Edge
OpenAI has launched gpt-realtime, a voice AI model for enterprises that delivers expressive, natural speech and follows complex instructions, targeting applications like customer support and real-time translation. The model operates on a speech-to-speech framework, enabling real-time vocal respon...
Read More » -
Liquid AI's LFM2-VL Model Brings Fast, Vision-Capable AI to Smartphones
Liquid AI has introduced LFM2-VL, a next-gen multimodal AI model optimized for smartphones and wearables, offering high speed and low resource usage while handling text and visual inputs. The model uses a unique Linear Input-Varying (LIV) approach and modular design, doubling GPU speeds and maint...
Read More » -
GPT-5 Now Free for All ChatGPT Users – OpenAI's Latest Release
OpenAI released GPT-5, its most advanced AI model, offering free access to ChatGPT users with specialized variants (Pro, mini, nano) for different needs, featuring fewer hallucinations and improved coding and sensitive query handling. GPT-5 introduces chain-of-thought reasoning for free-tier user...
Read More » -
OpenAI Targets Developers with GPT-4.1, A Fine-Tuned Coding Push
OpenAI has introduced a new family of AI models dubbed GPT-4.1, adding another layer to its already complex naming structure. This latest release includes three distinct versions – GPT-4.1, GPT-4.1 mini, and GPT-4.1 nano
Read More » -
Bard rebrands as Gemini
Meet Gemini, Google's groundbreaking AI. A leap from Bard, Gemini brings nuanced understanding and multimodal capabilities to your fingertips. Discover an AI that not only processes data but anticipates needs, setting a new standard in technological advancement. Gemini: Where innovation meets practicality.
Read More »