1. The Rise of Background Agents: Gemini Spark & Daily Brief
Google is moving away from AI that just chats, replacing it with AI that acts.
- Gemini Spark: This is a 24/7 personal AI agent that operates in the background, even when your laptop is closed. It connects directly to your Google Workspace (Gmail, Docs, Sheets) to handle complex digital chores autonomously—like compiling wedding RSVPs into a spreadsheet, building trip itineraries from flight confirmation emails, or managing your inbox. It operates under strict guardrails and asks for permission before taking high-stakes actions like spending money.
- Daily Brief: An out-of-the-box agent that organizes your day each morning, offering a personalized digest of your goals and suggesting priority next steps.
Gemini Spark, your new personal agent in the Gemini app that gets stuff done on your behalf, under your direction. It runs 24/7, integrates seamlessly with tools (starting with ours), and soon you can work with through chat and email. it is kind of Your very own corporate Open-Claw. gemini.google.com/spark
Examples
prompt1: Spreadsheet Creation and Formula Evaluation
Create a new Google Spreadsheet by putting a formula =GOOGLEFINANCE("CURRENCY:USDJPY") in cell “A1” of the first sheet. Then, get and show the value of cell “A1”.
Prompt 2: File Discovery in Folder Show the file list from a folder named “sample folder” in my Google Drive.
Prompt 3: Web Scraping and Document Generation
Fetch information from the URL https://tanaikech.github.io/about/, organize and summarize the details clearly into a new Google Doc, and finally display the URL of the generated Google Doc.
Prompt 4: Autonomous Gmail Event Triggering
When a new email is received from tanaike@hotmail.com, send the data to “Sheet1” in the Google Spreadsheet named Sample for Gemini Spark.
2. Next-Generation Models: Gemini 3.5 & Omni
The models powering Google’s AI have received massive upgrades, focusing on action-taking and media creation.
-
Gemini 3.5 Flash: Launched globally as the default model across many Google services, this model is specifically built to handle complex, multi-step agentic workflows and coding tasks at lightning speed. Gemini 3.5: the latest family of models–starting with Gemini 3.5 Flash. Everyone wants agents, MCP, A2A, and cross-cloud workflows, but identity and authorization are still the part that can make or break the whole experience.
-
Gemini Omni: A powerful new multimodal AI that blends text, photos, audio, and video. It allows users to generate and edit high-quality cinematic videos as easily as having a conversation. You can even create custom AI avatars of yourself to drop into the action. a new step in world models that can create anything from any input - starting with video. Nano Banana on steroids
Gemini Omni Conversational Video Editing: Refine and edit videos using natural language, modifying styles, actions, or camera angles while maintaining continuity.
Gemini Omni Multimodal Referencing: Combine inputs like images, text, and video to maintain precise control and consistency over your scene.
Gemini Omni Real-world Knowledge: Draws on Gemini’s extensive knowledge of history, biology, and narrative logic to construct compelling videos. https://deepmind.google/models/gemini-omni/ https://ai.google.dev/gemini-api/docs/omni#prompt-guide
3. The Rebuilding of Google Search
Google Search has received its biggest upgrade in 25 years with the introduction of AI Mode.
- Powered by the advanced reasoning of Gemini 3.1 Pro and Gemini 3 Pro, AI Mode uses a “Query Fan-Out” technique to break complex questions into subtopics and search multiple sources simultaneously.
- Instead of standard blue links, Search now dynamically builds Generative UIs—custom interactive layouts, dashboards, and visual timelines tailored exactly to your question.
- It also features Information Agents that monitor the web on your behalf 24/7, sending you detailed updates and actionable links on topics you care about.
4. Deep Ecosystem Integration
AI is now the connective tissue across all of Google’s hardware and software.
- Android 17 & Pixel: Launched in June 2026, Android 17 brings deep AI integration directly to mobile, including AI-assisted media creation, floating Bubbles for multitasking, and a completely redesigned “Neural Expressive” interface in the Gemini app that uses fluid animations and haptic feedback.
- Chrome AI: Chrome is transitioning into an AI workspace featuring Auto Browse, a Gemini-powered tool that can perform multi-step browsing tasks (like comparing products) while asking for confirmation before taking sensitive actions.
- Gemini Live: The Gemini app now allows you to switch seamlessly between talking and typing, making natural, real-time conversational AI a reality on the go.
Lyria
Create music Try a template or describe a track in chat. Create with Lyria 3.
codewiki
voice conversations
Use the Gemini Live API to give your app a voice and make your own conversational experiences.
Gemini intelligence
Embed Gemini in your app to complete all sorts of tasks - analyze content, make edits, and more
Pomelli by Google Labs for content
https://labs.google.com/u/0/pomelli/campaigns
Flow, AI filmmaking tool
https://labs.google/fx/tools/flow
gems
https://gemini.google/overview/gems/
stitch
https://stitch.withgoogle.com/
gemini
gemini.google.com
gemini cli deprecated now
antigravity
https://antigravity.google/product/antigravity-cli
firebase
http://firebase.google.com/ firebase studio https://console.firebase.google.com/
nano banana
https://deepmind.google/models/gemini-image/flash-lite/ nano banana 2 lite https://ai.google.dev/gemini-api/docs/image-generation#prompt-guide nano banana image editing mode - if you need reference check build with nano banana and all,
veo3
veo3 video editing mode,