How Google autocomplete works (and what it is not)
Where search predictions come from, why they change by country and over time, why some prefixes return nothing, and how GuessTheSearch turns that into a daily quiz.
By The GuessTheSearch team · · 7 min read
Start typing into a search box and a list drops down before you finish: the engineâs guess at what you are about to ask. GuessTheSearch is built entirely on that list. Every day we show you the start of a question and ask you to guess the top prediction. To play well â and to read the answers sensibly â it helps to know where those predictions come from.
This article sticks to what Google has said publicly about autocomplete, plus what we have observed from capturing thousands of predictions ourselves. We are an independent game and are not affiliated with Google, so we will not pretend to know anything about its systems beyond that.
Predictions, not suggestions
Google describes autocomplete results as predictions rather than suggestions, and the distinction matters. The feature is not recommending what you should search for. It is trying to predict what you are most likely typing, so you can finish your search faster â which is especially useful on a phone keyboard, or when you are not sure how to spell something.
According to Googleâs own explanations, those predictions are drawn from searches that have actually been made on Google. The system looks at what people have typed that begins the same way you have begun, and offers the most likely continuations.
What goes into a prediction
Google has named several broad factors that shape which predictions you see and in what order:
- Real queries. Predictions reflect searches people have actually typed, not phrases written by an editor.
- Popularity. Common searches are more likely to appear than rare ones.
- Freshness. Searches that are trending â rising in interest right now â can appear even if they would not top an all-time list.
- Language and location. The same letters typed in a different language or country can produce very different predictions.
- Your own history. If you are signed in and have search history turned on, some predictions can reflect your past searches.
Popularity is the easiest to see in action. Here are two differently phrased questions â one about a country, one about its people â that we captured on the same day:
- 1why doesnât france have air conditioning
- 2why doesnât france use ac
- 3why doesnât france allow dna testing
- 4why doesnât france help haiti
- 5why doesnât france annex monaco
- 1why donât french people have air conditioning
- 2why donât french people use ac
- 3why donât french people work in august
- 4why donât french people buy chateaux
The phrasing is different but the top two predictions are identical, because the curiosity behind them is the same. The same prediction tops our captured lists for Germany, Italy and the UK: a great many English-speaking searchers want to know why so many European homes go without air conditioning. A popular question tends to surface whichever way you ask it.
What predictions are not
Because predictions look like sentences, it is easy to read too much into them.
- They are not answers. âWhy is X so Y?â being predicted says only that people ask it. It does not mean X is Y, and the question may rest on a false premise.
- They are not Googleâs opinion. They reflect what searchers type, filtered by Googleâs policies. They are not an editorial view about any person, place or group.
- They are not a poll. Someone who types a question may be curious, sceptical, joking or checking a claim they heard. Search volume measures interest, not belief.
- They are not complete. Google removes some predictions under published policies, and in some contexts shows none at all. We look at that in detail in autocomplete and bias.
Why the same prefix gives different results in different places
Location and language make a bigger difference than most people expect. When we first ran our capture from a server in Zagreb, the prompt âjapan vsâ came back full of Croatian football fixtures. That is perfectly sensible for a Croatian searcher and useless for a quiz.
So GuessTheSearch pins every request to US English, using the standard language and country parameters (hl=en and gl=us). That has two effects. First, everyone playing the same puzzle is guessing against the same data, wherever they happen to be. Second, the daily capture does not drift depending on where our server runs.
Why some prefixes return nothing
One of the first things we learned building the game is that autocomplete is fussier about wording than you might think. Plenty of perfectly natural questions return no predictions at all. In our testing, âwhy are nurses soâ came back empty, while âwhy is nursing soâ returned a full list. âWhy are lawyers alwaysâ returned nothing, but âwhy donât lawyersâ returned plenty:
- 1why donât lawyers represent themselves in court
- 2why donât lawyers go by doctor
- 3why donât lawyers get called doctor
- 4why donât lawyers use the title doctor
- 5why donât lawyers call you back
- 6why donât lawyers take credit cards
We can only describe the pattern, not explain it from the inside. Some empty results are probably just a lack of popular searches that begin that way. Others may reflect Googleâs policies, which allow it to show fewer or no predictions for certain kinds of prompts. Either way, it has shaped the game: the pairing of question shapes and themes in GuessTheSearch is measured, not chosen. We probed each shape against live autocomplete and kept only the combinations where Google actually answers. That is why jobs never get a âWhy are ⊠always ___?â question, and why some subjects are phrased the way they are.
Grammar matters too. Each subject carries its own verb, so we ask âwhy are printers soâ rather than âwhy is printer soâ. The mismatched version returns a much thinner result set, because far fewer people type it.
What we get back, and what we keep
The raw response is messier than the tidy dropdown you see in a browser. A request for one prefix can include near-misses that only look like matches, the bare prefix with nothing after it, and pointers to forums and video sites. Before anything reaches a puzzle, our code:
- checks that the prediction starts with the exact question shape we asked for;
- checks that the subject sits directly before the completion;
- drops answers too long to fit on a button;
- removes platform noise, name collisions and anything our content filters reject;
- keeps the list visually distinct, so two options never differ by a single word.
The order of what survives is Googleâs order. The #1 prediction that remains is the correct answer; the others become decoys on the harder difficulty levels. Our editorial standards describe the content filters, and every recapture is reviewed by a person before it ships.
Predictions change over time
Because freshness is one of the factors, autocomplete is a moving target. Some predictions are evergreen â âwhy do Greeksâ is topped by breaking plates, a question with no news hook at all. Others are visibly tied to the moment they were captured:
- 1will ram ever go back down
- 2will ram ever come down in price
- 3will ram ever get her horn back
- 4will ram ever be cheap again
Three of those four are about memory prices, which says a lot about when they were captured. (The odd one out is about a character from an animated series; name collisions like that are a constant feature of autocomplete, and part of the charm.) Our captured technology questions are full of the same kind of time stamps: whether a particular operating-system version is out yet, whether an older one is still supported, whether programmers are being replaced by AI. Ask the same questions a year from now and many of them will have moved on.
How that becomes four different quizzes
That drift is exactly why GuessTheSearch offers four time periods. We capture predictions every day and keep the snapshots, so each period reads from a different slice of history:
| Period | What it plays against |
|---|---|
| Last week | A consensus of the last seven daily snapshots, so one noisy day cannot define the week |
| Last month | The snapshot nearest to thirty days ago |
| Last year | The snapshot nearest to a year ago |
| All time | The earliest snapshot we have on record |
The weekly consensus uses a simple ranked-vote method: each dayâs list votes for its completions by position, and the totals decide the weekâs order. Where our archive does not reach back far enough yet, a period falls back to our offline answer set, rotated so that the four periods are still genuinely different puzzles. As the archive grows, those older periods fill in with real snapshots automatically.
Playing the same theme across periods is one of the more interesting things you can do in the game. Shifts in the top answer are small, real records of what people were curious about at the time.
What this means when you play
- Think like a searcher. The top prediction is what many people type, not the cleverest or truest answer. Our strategy guide goes into this in depth.
- Think American English. The data is pinned to US English, so US spellings, sports and news cycles show up.
- Think about the moment. In the âLast weekâ period, recent events can push a fresh question to the top.
- Read the answers as questions. A prediction about a nationality or a star sign is a record of curiosity, not a verdict. Our posts on star signs, generations and European stereotypes show what that looks like in practice, and where stereotypes come from covers the psychology behind it.
The full rules, including how difficulty levels choose decoys, are on how to play.