EntitySEO.wikiThe Entity SEO Reference

Neural IR, BERT, MUM and passage ranking: patents and papers

From EntitySEO.wiki, the entity SEO reference · Last reviewed · By · Published by Local Blitz · How we research

Neural information retrieval (neural IR) uses deep learning models to understand queries and passages and to match them by meaning. Google has confirmed using BERT, MUM for some tasks, RankBrain, neural matching and passage ranking. This page lists the verified research behind those systems (the Transformer, BERT and T5) plus key retrieval papers and three Google patents.

Read this first: a patent shows that a company sought legal protection for a method. It does not show that the method is used in Google Search, used as written, or still used. Only systems Google itself names (for example in its ranking systems guide) are confirmed. Dates and legal status come from Google Patents, which notes that its legal status is “an assumption and is not a legal conclusion”. See how to read a patent.

Which neural systems has Google confirmed?

SystemWhat Google says
BERTApplied to ranking and featured snippets from 2019; Google said it would help Search “better understand one in 10 searches in the U.S. in English”.[1]
MUM“Uses the T5 text-to-text framework and is 1,000 times more powerful than BERT”[2]; “not currently used for general ranking”.[3]
Passage rankingIdentifies individual sections or “passages” of a page to judge relevance.[3]
RankBrain, neural matchingRelate words to concepts and match concepts in queries and pages (see query understanding).

Only the BERT and T5 papers below are directly tied to named systems by Google. The rest show how the field works. Officially documented

Which patents describe answer passages and Transformers?

Scoring candidate answer passages

Patent US 9,940,367 B1 Patent: use unconfirmed · Assignee Google LLC · Inventors Steven D. Baker, Srinivasan Venkatachary, Robert Andrew Brennan, Per Bjornsson, Yi Liu, Hadar Shemtov, Massimiliano Ciaramita, Ioannis Tsochantaridis · Priority Aug 13, 2014 · Filed Aug 12, 2015 · Granted Apr 10, 2018 · Status Active

What it describes: Scoring candidate answer passages for question queries. For each passage from responsive resources, the system computes a query-term match score and an answer-term match score, among other features, and selects an answer passage.

Why it matters for SEO: Passages that restate the question’s terms and contain the terms typical of good answers score well in this design. Put a direct, self-contained answer under each question heading. Practitioner practice

Context scoring adjustments for answer passages

Patent US 9,959,315 B1 Patent: use unconfirmed · Assignee Google LLC · Inventors Nitin Gupta, Srinivasan Venkatachary, Lingkun Chu, Steven D. Baker · Priority Jan 31, 2014 · Filed Jan 31, 2014 · Granted May 1, 2018 · Status Active

What it describes: Adjusting answer-passage scores using context: a heading vector describes the path through the page’s heading hierarchy from the root heading to the heading the passage sits under.

Why it matters for SEO: Your heading structure is context for every passage. Use a clean H1>H2>H3 hierarchy where headings describe their sections. AEO.wiki covers this in content structure. Practitioner practice

Attention-based sequence transduction neural networks

Patent US 10,452,978 B2 Patent: use unconfirmed · Assignee Google LLC · Inventors Noam M. Shazeer, Aidan Nicholas Gomez, Lukasz Mieczyslaw Kaiser, Jakob D. Uszkoreit, Llion Owen Jones, Niki J. Parmar, Illia Polosukhin, Ashish Teku Vaswani · Priority May 23, 2017 · Filed Jun 28, 2018 · Granted Oct 22, 2019 · Status Active

What it describes: Attention-based sequence transduction neural networks: the encoder-decoder architecture built from attention layers, known as the Transformer.

Why it matters for SEO: The Transformer is the architecture behind BERT, T5 and today’s large language models; this is Google’s patent on it.

Which research papers underpin BERT, MUM and neural retrieval?

Attention Is All You Need

Authors Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit et al. · Published NeurIPS 2017 (2017) · arXiv:1706.03762 Research paper

What it describes: The Transformer: a model built entirely on attention, which weighs how every word relates to every other word.

Why it matters for SEO: Understanding words in full context is why modern search handles long, conversational queries.

BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Authors Jacob Devlin, Ming-Wei Chang, Kenton Lee, Kristina Toutanova · Published NAACL-HLT 2019 (2019) · ACL Anthology Research paper

What it describes: BERT, a Transformer pre-trained to read text in both directions, then fine-tuned for tasks such as question answering.

Why it matters for SEO: Google applied BERT to Search in 2019.[1] Small words like “to” and “for” now change meaning, so write naturally and precisely. Officially documented

Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer

Authors Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee et al. · Published Journal of Machine Learning Research 21 (2020) · JMLR Research paper

What it describes: T5, which casts every language task as text in, text out.

Why it matters for SEO: Google says MUM uses the T5 text-to-text framework.[2] Officially documented

MS MARCO: A Human Generated MAchine Reading COmprehension Dataset

Authors Payal Bajaj, Daniel Campos, Nick Craswell, Li Deng et al. · Published arXiv preprint (2016) · arXiv:1611.09268 Research paper

What it describes: A large Microsoft dataset of real Bing queries with human-written answers and relevant passages.

Why it matters for SEO: The standard benchmark for passage ranking; much of the research below was measured on it.

Natural Questions: A Benchmark for Question Answering Research

Authors Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti et al. · Published Transactions of the Association for Computational Linguistics (2019) · MIT Press Research paper

What it describes: A Google dataset of real Google queries paired with Wikipedia pages annotated with long and short answers.

Why it matters for SEO: Answers are often a paragraph plus a short fact inside it: the same shape as an answer-first paragraph.

Passage Re-ranking with BERT

Authors Rodrigo Nogueira, Kyunghyun Cho · Published arXiv preprint (2019) · arXiv:1901.04085 Research paper

What it describes: Using BERT to re-rank candidate passages for a query, which sharply improved results on MS MARCO.

Why it matters for SEO: Shows the pattern of cheap retrieval followed by a smarter re-ranking step. Each passage is judged on its own.

ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERT

Authors Omar Khattab, Matei Zaharia · Published SIGIR 2020 (2020) · arXiv:2004.12832 Research paper

What it describes: A retrieval model that keeps BERT-level quality while being fast enough for large collections.

Why it matters for SEO: Makes passage-level neural retrieval practical at scale.

Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Authors Nils Reimers, Iryna Gurevych · Published EMNLP-IJCNLP 2019 (2019) · ACL Anthology Research paper

What it describes: Producing sentence embeddings that can be compared by similarity, the basis of many semantic search tools.

Why it matters for SEO: Many SEO content tools use embeddings like these to measure topical similarity.

Dense Passage Retrieval for Open-Domain Question Answering

Authors Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov et al. · Published EMNLP 2020 (2020) · ACL Anthology Research paper

What it describes: Retrieving passages with learned dense vectors instead of word matching, outperforming BM25 on several QA benchmarks.

Why it matters for SEO: Passages can be found by meaning even without shared keywords, so clarity beats keyword stuffing.

Large Dual Encoders Are Generalizable Retrievers

Authors Jianmo Ni, Chen Qu, Jing Lu, Zhuyun Dai et al. · Published arXiv preprint (2021) · arXiv:2112.07899 Research paper

What it describes: Google research showing that scaling up dual-encoder retrievers (GTR) improves how well they work on new domains.

Why it matters for SEO: Dense retrieval keeps improving, which rewards clearly written, self-contained passages.

What should you do with this?

Frequently asked questions

Did BERT replace RankBrain?

No. Google lists both BERT and RankBrain as separate AI systems in its ranking systems guide.

Is MUM used for ranking?

Google's ranking systems guide says MUM is not currently used for general ranking, only for specific applications such as improving featured snippet callouts.

What is passage ranking?

An AI system Google uses to identify individual sections or passages of a page to better understand how relevant the page is to a search.

References

Pages accessed September 29, 2026 unless a date is given. See all sources and our editorial policy.

  1. ^ "Understanding searches better than ever before (Pandu Nayak)". Google (The Keyword). Published October 25, 2019.
  2. ^ "MUM: A new AI milestone for understanding information (Pandu Nayak)". Google (The Keyword). Published May 18, 2021.
  3. ^ "A guide to Google Search ranking systems". Google Search Central.