// AI NEWS · FEATURE

Google Search Console Adds a "Multimodal" Filter — Image-Based Searches From Lens and Circle to Search Now Show Up in Your Performance Report

Rolling out globally from September 24, 2026, the Search type filter now splits "Web" into Text-based and Multimodal, so you can finally see impressions and clicks from visual searches — Lens, Circle to Search, camera and screenshot lookups. There is one catch: no query data.

2026-09-25 · by Moe Ameen

What happened

On September 24, 2026, Google began rolling out globally a new Multimodal filter in Search Console's Performance report. Opening the Search type filter now splits the "Web" option into two: Text-based and Multimodal. Per Google's documentation, "Web-multimodal tracks search results triggered by a query that uses an image, photo, or screenshot. Text-only queries are tracked as Web text-based." In other words, the searches where someone pointed a camera, uploaded a screenshot, or circled something on their screen are now broken out from ordinary typed queries.

The multimodal bucket covers searches made with Google Lens, Circle to Search on Android, image uploads to Google Search, and Chrome's right-click "Search this image." Google product managers Harsh Kharbanda and Moshe Samet framed the change as visibility into a discovery surface that was previously invisible: "This update is designed to give you insights into how your content is surfaced when users search using images (such as with a smartphone camera)." The split appears in the standard Search results performance report and also feeds the generative AI performance report.

There is a real limitation to understand before reading anything into the numbers. You get impressions, clicks, and position for multimodal traffic, sliceable by page, country, and device — but you do not get the queries. Google's stated reason: "Because multimodal searches mostly use images rather than text, specific text query data isn't available for this traffic." So the classic keyword-to-content loop that text search enables does not work here. The generative AI report, separately, still lacks click and query data of its own. Treat availability as a rollout-window snapshot and confirm what has landed in your own Search Console before drawing conclusions from a single week.

Why it matters for creators

  • Visual discovery is now measurable, which means it is now a channel you can manage. For the first time you can see whether people are finding your pages by pointing a camera or uploading an image, instead of guessing — the surface where "search" increasingly means Lens and Circle to Search rather than a text box.
  • Image-heavy niches should look first. Products, food, interiors, fashion, plants, parts, landmarks, anything people photograph to identify or shop — these are exactly the queries multimodal search serves, so a spike in the Multimodal filter is a signal to produce more strong, on-brand visual content for that topic.
  • No query data changes the optimization loop. You cannot reverse-engineer keywords from visual searches, so the lever is not keyword targeting — it is producing a high volume of clear, well-labeled, distinctive images that a visual engine can match and surface, then watching which pages earn multimodal impressions.
  • It rewards owning original imagery. Lens and Circle to Search match against the actual pixels on your page, so generic stock and templated graphics compete poorly against distinctive, first-party visuals — the creators who produce their own consistent image sets have the edge here.
  • The metric is real, but it is a leading indicator, not a payout. A multimodal impression is exposure inside a visual result, not a guaranteed visit — the durable response is producing enough on-brand visual content to show up across many image searches, not chasing a single one.

How to act on this with Kompozy

This filter quietly confirms a shift most content strategies still ignore: a growing share of discovery is visual, not textual. People aim a camera at a product, screenshot a post, or circle an object on screen — and Google matches that against the actual images on your pages. The uncomfortable implication is that being findable in multimodal search is a production problem, not a keyword problem. There is nothing to optimize for a phrase, because there is no phrase; there is only whether you have published enough strong, distinctive, on-brand imagery for a visual engine to surface. That is the exact ceiling [Kompozy](/) is built to lift. It is not a text-first tool bolted onto images — it generates visual content as a first-class output: scene photo posts, infographic posters, face-locked [Persona Photos](/glossary/output-buckets), brand-exact [Carousels](/glossary/hyperframes), quote graphics, and short video, all in your look, then [Autopilot](/glossary/autopilot) schedules and publishes them across the eight social platforms plus blog and email behind a per-post review gate.

The angle that fits this specific news is visual volume with brand consistency. Because multimodal search returns no query data, you cannot tune toward keywords — the winning move is to produce many clear, distinctive, first-party images across your topics so a visual match has something to find, and to keep every one of them recognizably yours so exposure compounds into a brand people recall. A [Persona Brief](/glossary/persona-brief) plus HyperFrames hold that look across the whole set, which is what separates a match-worthy original image from interchangeable stock. Then the new Multimodal filter becomes your feedback loop: watch which pages earn image-driven impressions and make more of what works. The honest boundary — Kompozy cannot make Lens rank you, and no tool can fake a distinctive visual; matches are earned by publishing genuinely good, original imagery. What it removes is the production ceiling that keeps most creators from ever making enough visual content to be found this way. For the deeper playbook, see [image SEO for AI-powered search](/guides/image-seo-for-ai-powered-search), [AI search impressions in Google](/guides/ai-search-impressions-in-google), and [how to set up AI search performance reporting in Search Console](/how-to/set-up-ai-search-performance-reporting-in-search-console).

Quick takeaways

  • On September 24, 2026, Google began rolling out globally a Multimodal filter in Search Console — the Search type filter now splits "Web" into Text-based and Multimodal.
  • Multimodal covers image-based searches: Google Lens, Circle to Search on Android, image uploads to Google Search, and Chrome's "Search this image."
  • You get impressions, clicks, and position by page, country, and device — but no query data, because these searches use images rather than text.
  • The split appears in both the Search results report and the generative AI performance report (which itself still lacks clicks and queries).
  • The takeaway for creators is visual volume: produce a lot of distinctive, on-brand first-party imagery so a visual engine has something to match — exactly what Kompozy generates and publishes.

Frequently asked questions

What is the Multimodal filter in Google Search Console?

It is a new option in the Search Console Performance report, rolling out globally from September 24, 2026, that splits the "Web" search type into Text-based and Multimodal. Google defines it as tracking "search results triggered by a query that uses an image, photo, or screenshot," while text-only queries stay under Web text-based. It covers Google Lens, Circle to Search on Android, image uploads to Google Search, and Chrome's "Search this image."

Why does the Multimodal filter have no query data?

Because these searches are made with images rather than typed text. Google's stated reason is that "because multimodal searches mostly use images rather than text, specific text query data isn't available for this traffic." You can see impressions, clicks, and position for multimodal traffic, sliceable by page, country, and device — but not the queries that produced them, so the usual keyword-to-content loop does not apply.

How do I optimize for multimodal and visual search?

Since there are no keywords to target, the lever is producing a high volume of clear, distinctive, on-brand first-party images that a visual engine like Lens can match against your pages, and then watching which pages earn multimodal impressions in the new filter. Original, consistent imagery beats generic stock or templated graphics. A content engine like Kompozy generates that visual set — photo posts, carousels, infographics, persona images, and short video — under one brand brief and publishes it across platforms.

Does the Multimodal filter apply to the generative AI performance report?

Yes. The Text-based and Multimodal split appears in both the standard Search results performance report and the generative AI performance report. Note that the generative AI report separately still lacks click and query data of its own, so read the multimodal split there as visibility into which content surfaces, not as a complete performance picture.

Related news

← All AI news · Get started →