Book a demo

Cross-Modal Retrieval Patents: Leaders & Filing Trends 2026

Cross-Modal Retrieval Patents: Leaders & Filing Trends 2026
https://www.patsnap.com/resources/blog/rd-blog/vector-databases-and-retrieval-cross-modal-vector-retrieval-patent-landscape-patent-landscape/ · Patsnap · data cut-off 2026-07-31 · downloaded from the live page
Patent Landscape · Cross-Modal Retrieval
Cross-modal retrieval patents: who is building the claim base for multimodal search
  • Filings quadrupled in three years. Documented cross-modal and multimodal retrieval filings grew from 72 in 2021 to 351 in 2024, a +388% rise over that span — and 2024 is the last year the trend can be read as complete.
  • The field is still open at the top. The leading assignee holds 139 records out of 2,230, and the top 5 combined account for just 14.7% of all records in scope — no single company has locked down the claim space.
  • G06F and G06N dominate, but coverage past them thins fast. 75.7% of records sit in G06F and 40.2% in G06N, while healthcare informatics (G16H, 5.2%) and speech/audio (G10L, 4.8%) carry far fewer filings relative to the core.
Get a prior-art report on your approach
2,230
Published Records
15%
Top-5 Share of All Records
+388%
Filing Growth 2021→2024
CN
Leading Jurisdiction

Filing growth compares 2021 (72 records) with 2024 (351) — a three-year span. 2024 is the most recent year we treat as complete: publication lags filing by roughly 18 months, so 2025 onwards are still filling in and any growth rate that ends there would understate the field. Top-5 share is the combined record count of the five largest assignees divided by all 2,230 records in scope (CR5), not by the ranked leaders only.

Published byPatsnap Research··7 min readSourced from Patsnap Eureka
Overview

What this landscape covers

This landscape tracks patent activity around cross-modal and multimodal retrieval — systems that map images, text, audio or video into a shared embedding space so that a query in one modality can retrieve results in another. The search spans cross modal retrieval, multimodal embedding, image-text retrieval and related terms across 2,230 records published between 2015 and 2026.

Filing activity is concentrated in electric digital data processing and AI-model computing classes, with smaller but distinct pockets in commerce, image recognition, healthcare informatics and speech processing. The assignee base is led by a handful of large technology filers, but the top 10 combined still cover under a fifth of all records, leaving substantial room for new entrants to stake out specific application claims.

Filing activity by year and IPC class
  1. 1GOOGLE LLC139
  2. 2MICROSOFT TECHNOLOGY LICENSING LLC63
  3. 3ADOBE INC47
  4. 4SAMSUNG ELECTRONICS CO LTD42
  5. 5BEIJING BAIDU NETCOM SCI & TECH CO LTD37
  6. 6BLANDING HOVENWEEP LLC32
  7. 7VELLORE INSITUTE OF TECH23
  8. 8THE INVENTION SCIENCE FUND 1 LLC18
  9. 9SEARETE LLC16
  10. 10DELL PROD LP13
Source: Patsnap Eureka. Assignee ranking and totals. Derived from a Patsnap search on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish.Run this in Eureka MCP

Let an AI agent run this analysis on your own technology

Pick a task. Every answer cites the patents behind it.

10,000 free credits to start
The data

Filing trend and technology composition

The two views below show how fast the field is moving and where filings actually sit once broken out by IPC subclass.

Filings climbed sharply from 2021 onward

Annual filings rose from 30 in 2017 to a documented peak of 677 in 2025, with 2021-to-2024 alone showing +388% growth (72 to 351). 2025 and 2026 figures are still filling in as publications lag behind filing dates by roughly 18 months, so the most recent two years should not be read as a slowdown.

Filings climbed sharply from 2021 onward0200400600800302017201820192020202120222023202467720253222026Most recent year is partial — publication lag means later filings are not yet visible.

Filings cluster in two core classes, then fragment

G06F (75.7% of records) and G06N (40.2%) anchor the field as the general data-processing and AI-model substrate. Below that, G06Q (13.3%), G06V (10.0%) and G06T (8.4%) mark distinct application layers, while G16H, G10L and G06K each sit under 6% — evidence of real but comparatively thin coverage in healthcare, speech and recognition-specific claims.

Filings cluster in two core classes, then fragmentG06F · Electric digital data processi…1,68875.7%G06N · Computing based on AI models89640.2%G06Q · Business, commerce & admin dat…29713.3%G06V · Image/video recognition22310.0%G06T · Image data processing & genera…1888.4%G16H · Healthcare informatics1175.2%G10L · Speech & audio analysis/synthe…1084.8%G06K · Data recognition & presentation954.3%Other50622.7%

Shares are the percentage of the 2,230 records in scope. A patent can carry several IPC classes, so the shares add up to more than 100%.

Source: Patsnap Eureka. Filing trend and technology composition. Derived from a Patsnap search on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish.

Go deeper on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape with Eureka

This page is one run against one query. Ask Eureka your own question about vector databases & retrieval — cross-modal vector retrieval patent landscape and every answer comes back with the patent numbers behind it.

Try Eureka
Representative filing

A recent claim shows where the field is heading

Filed 2026-05-28 · Google LLC
US20260147800A12026-05-28

Multimodal Retrieval Augmented Visual Question Answering

GOOGLE LLC

Systems and methods for visual question answering that obtain a multimodal input, generate a multimodal embedding, perform a search based on that embedding, determine a relevant passage from the results, and process the input and passage with a generative model to produce a response. The generative model is described as configured to orchestrate multimodal searches, rank results, determine relevant passages and generate the final response.This pattern — embedding generation feeding a search step that feeds a generative model — is the structure most recent filings in this space follow.

US20260147800A1 — patent drawing 1US20260147800A1 — patent drawing 2
View full record
Most-cited records in this landscape
#Publication no.Patent titleCitations
1US6850252B1Intelligent electronic appliance system and method4,059
2US6400996B1Adaptive pattern recognition based control system and method2,342
3US6640145B2Media recording device with packet data interface1,591
4US20070053513A1Intelligent electronic appliance system and method1,452
5US7006881B1Media recording device with remote graphic user interface1,218
6US7813822B1Intelligent electronic appliance system and method1,198
7US6510406B1Inverse inference engine for high performance web search720
8US20060200253A1Internet appliance system and method558
9US20020151992A1Media recording device with packet data interface501
10US6862710B1Internet navigation using soft hyperlinks493

Citation counts reward older filings that have had more time to accumulate references — read them as a signal of historical influence, not current relevance.

Each row carries its publication number; clicking a row searches Eureka by that number.

Source: Patsnap Eureka. Citation counts and representative records. Derived from a Patsnap search on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish.Run this in Eureka MCP
Run it yourself

Put your own technology through the same analysis

 
Where to run it
Fastest

Eureka on the web

When you want the answer in the next five minutes.

The agent works the prompt against patents and technical literature, citing every source.

Run your analysis now →
For builders

MCP server & REST API

When it has to run inside your own pipeline.

Patent search, landscape analysis and assignee resolution as MCP tools. Drop them into any agent framework, or call REST directly.

Browse MCP servers →
Insights

What the numbers mean for filing strategy

Three patterns stand out once the raw counts are broken down by year, class and assignee.

Growth
+388%
2021 → 2024 filings

The field's growth curve is real and recent

Filings went from 72 in 2021 to 351 in 2024. That is a genuine acceleration, not an artefact of a longer search window, and it means prior art from even three years ago may already be outdated by claim scope filed since.

2021–2024, complete years only
Concentration
14.7%
top 5 share of 2,230 records

No single filer controls the space

The leading assignee holds 139 records; the top 5 combined hold 328, or 14.7% of all 2,230 records in scope. That leaves the large majority of filings spread across a long tail of single- and few-filing entrants.

Top 5 of 100 ranked assignees
Technology mix
75.7%
records in G06F

Core infrastructure claims dominate, application claims trail

Three in four records touch G06F (electric digital data processing) and two in five touch G06N (AI-model computing). Healthcare informatics, speech processing and recognition-specific classes each sit under 11% — these are the branches with proportionally lighter claim density.

Shares sum above 100% — records carry multiple IPC classes
Geography
832
China filings (receiving office)

China and the US lead filing volume, India is a fast-growing third

China's receiving office recorded 832 filings and the US 759, with India at 283 and WIPO PCT filings at 167. Any freedom-to-operate check for this space needs to clear at least these three jurisdictions before Europe or Australia.

By receiving office, all years
Eureka AI Agent
Looking for what nobody has claimed yet?

Eureka can read the same corpus for gaps instead of for coverage: under-claimed branches adjacent to vector databases & retrieval — cross-modal vector retrieval patent landscape, with the prior art for and against each one.

Find the white space →
Source: Patsnap Eureka. Co-assignee relationships and derived observations. Derived from a Patsnap search on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish.Run this in Eureka MCP
Players

Who is filing, and who is accelerating

The leaderboard mixes large diversified technology filers with at least one fast-moving academic institution — and several of the historically largest filers have gone quiet in the most recent year.

Leader by volume
139
records

One filer leads by a wide margin over fifth place

The top-ranked assignee holds 139 records against 37 at fifth place and 13 at tenth — a steep drop-off that suggests the leader's position rests on sustained, broad filing rather than a single concentrated push.

Leader vs. 5th vs. 10th place
Momentum
+129%
YoY, latest year

An academic filer is growing fastest right now

Vellore Institute of Technology filed 16 records in the latest year, up 129% year-on-year — the only tracked assignee still accelerating while several large corporate filers show year-on-year declines of 80% or more.

Latest year vs. prior year
Co-filing
21
strongest co-assignee pair

Co-assignment is rare and concentrated in one pairing

Only 10 co-assignee pairs appear across the dataset, and the strongest single pairing accounts for 21 joint records — far ahead of any other pair, which sit at 2. Collaborative filing is not yet a common strategy in this field.

10 pairs total, all years
🔍
Under-claimed sub-areas worth a closer look
Filing density outside the core classes is comparatively thin — these are starting points for a freedom-to-operate scan, not confirmed open space.
Cross-modal retrieval for clinical imaging recordsSpeech-to-visual embedding alignmentMultimodal retrieval for e-commerce catalog searchRecognition-layer fusion for video-text matchingEmbedding compression for on-device multimodal search
Rank all filers by momentum →
Recent-year filing momentum by assignee
AssigneeRecent yearYoY
VELLORE INSITUTE OF TECH16+129%
Google LLC6-81%
Microsoft Technology Licensing, LLC1-89%
Adobe Inc.0-100%
Samsung Electronics Co., Ltd. (Korea)0-100%
Beijing Baidu Netcom Science & Technology Co., Ltd.0-100%
Searete LLC0
HOFFBERG STEVEN M0
Source: Patsnap Eureka. Assignee-level momentum. Derived from a Patsnap search on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish.Run this in Eureka MCP
What's next

Where to take this

The dataset points to a field still forming its claim boundaries rather than one settling into consolidation.

Check freedom to operate before filing in core classes

With 75.7% of records in G06F and 40.2% in G06N, any new filing that touches general embedding generation or retrieval architecture needs a targeted prior-art search in those two classes first.

Run a novelty check in Eureka →

Watch the academic filer accelerating against the trend

Vellore Institute of Technology's 129% year-on-year growth stands out against declining large-corporate filers — worth tracking for licensing or collaboration signals before the position solidifies.

Track assignee activity in Eureka →

Scope claims toward the thinner application classes

Healthcare informatics, speech/audio and recognition-specific classes each carry under 11% of records — a narrower but potentially clearer path to allowable claims than the crowded G06F/G06N core.

Explore white space in Eureka →
Source: Patsnap Eureka. Forward-looking reading of the same dataset. Derived from a Patsnap search on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish.Run this in Eureka MCP
FAQ

Common questions about this landscape

Answers are grounded in the same dataset. Derived from a Patsnap search on Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish.Run this in Eureka MCP

Research Vector Databases & Retrieval — Cross-Modal Vector Retrieval Patent Landscape in depth with Eureka

Go past this page: query the whole vector databases & retrieval — cross-modal vector retrieval patent landscape corpus yourself, in your own scope.
Every answer comes back with patent numbers you can open.

Try Eureka

Disclaimer. This page is generated from Patsnap Eureka data drawn from a limited snapshot of global patent and scientific-literature records, and is provided for general information and reference only.

Patent data carries inherent limitations: recent filings (typically the most recent 18–24 months) are under-counted due to standard publication lag; counts may be reported at either a patent-family or a patent-record basis and are not always directly comparable; classification, applicant-name, and citation data may contain errors, duplicates, or omissions; and the underlying search query defines and constrains the scope shown. As a result, the analysis may be incomplete or inaccurate and may not reflect the full technology landscape.

Nothing on this page constitutes an exhaustive prior-art, novelty, freedom-to-operate, or validity search, nor does it constitute legal, financial, investment, or professional advice, and it should not be relied upon as such. Any patent, commercial, or strategic decision should be verified independently and reviewed with qualified patent, legal, and domain professionals. Patsnap makes no warranties, express or implied, as to the accuracy, completeness, or fitness for any particular purpose of the information presented.

Machine translation. Assignee and organisation names originally recorded in Chinese, Japanese or Korean have been rendered into English by an AI translation step so that the tables stay readable. These renderings are best-effort and may not match a company’s registered English name; the original name is what the underlying patent record carries, and it is what any Eureka query launched from this page uses.

Help us improve this page

Found incorrect or outdated information? Let us know and we'll get it fixed.