
The results you see after typing a query into Google don’t appear by accident. Every URL that surfaces has been evaluated—directly or indirectly—by a group of people called Search Quality Raters. Their work shapes the algorithmic adjustments that determine whether a page ranks first or vanishes entirely. I’m Kyle Brennan, and I want to lay out exactly what these raters do, what their guidelines demand, and how their feedback loops into the ranking systems you deal with every day.
What Search Quality Raters Actually Do
A common misunderstanding is that raters change rankings directly. They don’t. A rater cannot move your site from position 11 to position 3. Instead, they evaluate a sample set of search results against a detailed handbook called the Search Quality Evaluator Guidelines—a 170-page document Google publishes publicly. Engineers use the evaluations to test proposed algorithm changes. If a rater marks a page as “low quality” and the algorithm currently ranks it high, that mismatch becomes a data point for refinement.
Raters work on specific tasks. They might compare two sets of search results side by side to determine which is better. Or they might assess a single page for how well it meets the needs of someone making a particular query. The work is granular. A rater notes whether a page has clear authorship, whether the content demonstrates expertise, and whether the page’s purpose is immediately obvious. These observations get aggregated into metrics that engineers use to run experiments.
Page Quality Ratings and E-E-A-T
The core of a rater’s job is assigning a Page Quality (PQ) rating. The scale runs from “Lowest” to “Highest,” and the difference often comes down to E-E-A-T: Experience, Expertise, Authoritativeness, and Trustworthiness. Google’s guidelines don’t weigh all pages the same way. A gossip blog and a medical information site are held to different standards. For a site that could affect someone’s financial stability or health—what the guidelines call “Your Money or Your Life” (YMYL) pages—the threshold for high quality is steep.
Here’s a concrete example. If a rater looks at a page offering tax advice, they check whether the author is a certified accountant or tax attorney, whether the information is current, and whether the site has a reasonable reputation. A forum post from an anonymous user with no cited sources would get a “Low” or “Lowest” rating. That rating doesn’t penalize the forum post immediately. But when thousands of similar ratings feed into a machine learning system, the pattern teaches the algorithm to devalue anonymous, unsourced financial content across the board.

How Raters Assess Needs Met
Beyond page quality, raters evaluate something called “Needs Met.” This measures how well a search result satisfies the intent behind a query. The scale runs from “Fully Meets” to “Fails to Meet.” A query like “weather in Phoenix” has a straightforward intent—the user wants a local forecast. If the top result is a page about the city’s climate history, a rater would mark it as failing to meet the immediate need. That feedback tells engineers that the algorithm is misinterpreting the query’s meaning.
Intent classification is where this gets precise. Raters distinguish between “Know” queries (seeking information), “Do” queries (seeking to perform an action, like buying or downloading), and “Visit” queries (seeking a specific website). A page that perfectly answers a “Know” query might be useless for a “Do” query. If someone searches “install a water heater” and the top result is a Wikipedia article on water heaters, a rater flags it because the user clearly wants a step-by-step guide, not an encyclopedia entry. Over time, this feedback trains the ranking system to match page type to query type more accurately.
The Mobile and Local Factors
Raters also consider the search context. A query performed on a mobile device in downtown Chicago at 6 p.m. has different implications than the same query typed on a desktop at noon. The guidelines instruct raters to note whether a result makes sense for a user on the go. For example, a mobile search for “pizza” should ideally return nearby restaurants with click-to-call buttons and hours, not just national chains’ corporate pages. Raters who consistently mark local, mobile-optimized results as meeting needs better than generic ones reinforce the importance of those signals in the algorithm.
The Feedback Loop Between Raters and Rankings
Here’s the sequence: engineers build a candidate algorithm update. They run it on a set of queries. Raters evaluate the old results and the new results without knowing which is which. If the new results score higher on average across Page Quality and Needs Met, the update may launch. If scores drop, the update is reworked. This is not a theoretical exercise. Every major search update—Panda, Penguin, the Helpful Content System—was shaped by rater evaluations before it reached users.
So what raters are trained to value eventually becomes what the algorithm values. If raters penalize pages with intrusive ads, algorithmic classifiers eventually learn to detect intrusive ad layouts. If raters reward pages with transparent author bios and original research, those features gain weight. The relationship is indirect but unmistakable. The guidelines are a lead indicator of where Google’s ranking priorities are headed.

What This Means for Site Owners
Reading the Search Quality Evaluator Guidelines is one of the most practical things you can do if you depend on organic search traffic. The document tells you exactly what raters look for. For any page, ask: Is the purpose clear? Is the content created by someone with demonstrable experience? For YMYL topics, can you cite original sources and point to a positive reputation? If you’re writing about a medical condition and you’re not a doctor, the guidelines explicitly state that you should include references to authoritative medical organizations. A rater will check for that.
Reputation research is another angle. Raters look for external signals about a site’s standing—reviews, news articles, references from experts. If your site is mentioned in a reputable publication or a professional directory, that matters. If the only references come from your own press releases, it doesn’t. The guidelines instruct raters to perform independent reputation searches. So if you haven’t checked what appears when someone searches your brand name plus “reviews” or “complaints,” you’re missing a factor that feeds directly into quality assessments.
FAQ
Do Search Quality Raters work directly for Google?
No. Google contracts with companies like Appen and Raterlabs, which hire raters as independent contractors. Raters operate under non-disclosure agreements and follow Google’s written guidelines, but they are not Google employees and have no direct access to ranking controls.
Can a rater’s evaluation get my site penalized?
Not directly. A single rater’s assessment never triggers a manual action or ranking drop. Rater data is used in aggregate to measure the performance of algorithm updates. However, if your site consistently exhibits the patterns that raters are trained to flag as low quality, those patterns may eventually be targeted algorithmically.
How often does Google update the rater guidelines?
Google updates the public guidelines roughly once a year. Changes often refine definitions of E-E-A-T, add examples, or clarify how to handle new content types like short-form video. Tracking the changelog can reveal shifts in what Google is prioritizing.
Are all pages rated on the same criteria?
No. The guidelines apply different standards depending on the page’s purpose and topic. A humor site is not expected to have the same level of formal expertise as a medical journal. Raters are explicitly told to adjust their expectations based on the page’s “beneficial purpose” and whether the topic falls into YMYL categories.