Data to 2026-07-31

Vision Language Model Architecture Patents: Filing Trends and White Space

48.4% of all 254 records sit with just five assignees, so the claim space at the architectural core is already crowded before you draft. Filings moved from 8 in 2021 to 98 in 2024, a +1125% increase over that span, before 2025 posted the observed peak of 100 records. Because publication lags filing by around eighteen months, 2025 and 2026 figures are still filling in and should not be read as the trend flattening or reversing.

A three-year filing surge, then a data lag
Complete Incomplete
A three-year filing surge, then a data lag02550751000201720182019202020212022202320241002025182026Most recent year is partial — publication lag means later filings are not yet visible.
254
Published Records
48%
Top-5 Share of All Records
+1125%
Filing Growth 2021→2024
US
Leading Jurisdiction
Top filers · published records
  1. 1NVIDIA CORP76
  2. 2GDM HOLDING LLC15
  3. 3GOOGLE LLC14
  4. 4ADOBE INC10
  5. 5SAMSUNG ELECTRONICS CO LTD8
Published by Patsnap Research·

See the full vision language model architectures analysis in Eureka

  • The complete ranking, not just the top five
  • Every IPC branch with its share of the corpus
  • The most-cited records, and where claim space is still thin
Read more about this analysis in Eureka
FAQ

Common questions on vision language model patents

Who holds the most patents in vision language model architectures?+

One assignee leads with 76 of the 254 records in scope, well ahead of the rest of the ranked list, where fifth place holds 8 and tenth place holds 4. The top five assignees combined account for 48.4% of all records, and the top ten reach 58.3%. That leaves a long tail of universities, startups and regional research institutes each holding a small handful of filings, so the field has one dominant filer plus a wide field of smaller players rather than a tight oligopoly.

Is the vision language model patent field growing or slowing down?+

It is growing sharply through the last complete data year. Filings rose from 8 in 2021 to 98 in 2024, a +1125% increase, with a further peak of 100 records in 2025. Figures for 2025 and 2026 will keep rising as publications catch up, since patent publication typically lags filing by about eighteen months, so any apparent slowdown in the newest years should not be read as the field cooling.

What patent classes cover vision language model technology?+

The dataset spans several IPC subclasses, with G06V (image and video recognition) at 44.5% of the 254 records and G06F (electric digital data processing) close behind at 44.1%. G06N, covering AI-model computing, appears in 38.6% of records. Because a single record can carry multiple classes, these shares add up to more than 100%; narrower application classes like robotics (B25J), healthcare informatics (G16H) and pictorial communication (H04N) each sit under 6%, marking them as thinner overlay categories rather than core claim territory.

Disclaimer. This analysis is based on Patsnap Eureka data drawn from a limited snapshot of global patent records and is provided for general information and reference only. Patent data carries inherent limitations — recent filings are under-counted because of publication lag, counts may be on a record or family basis, classification and applicant-name data may contain errors or duplicates, and the underlying search query defines the scope shown — so the analysis may be incomplete or inaccurate and may not reflect the full technology landscape.

Nothing here is an exhaustive prior-art, novelty, freedom-to-operate or validity search, nor does it constitute legal, financial or professional advice, and it should not be relied upon as such. Verify independently and review with qualified patent and legal professionals before acting on it.

Method: Filing trend and technology composition. Derived from a Patsnap search on Vision Language Model Architectures covering 2015–2026, data cut-off 2026-07-31. Counts reflect published records only and shift as new filings publish. Every share divides by all records in scope. Data: Patsnap Eureka. See the full landscape report.