Other Version: عربي
Home | Board of Directors | Media | Veterinary | News | References| Downloads | Contact Us

How Football Player Ratings Work: What the Score Out of Ten Actually Measures

A football player rating is a single score, usually on a scale of zero to ten, that summarises one player's performance in one match. Modern ratings are generated automatically from event data, the same records of passes, shots and tackles that appear on RubiScore match pages, and this explainer covers how they are built and what they miss.

Ratings are popular because they compress a match into one number that is easy to compare. That convenience is also their weakness. The event data behind every rating is the kind listed at https://rubiscore.com, and understanding how a rating turns those events into a score is the key to reading it sensibly.

From Journalists' Marks to Algorithms

Ratings out of ten began as a journalistic tradition. Newspapers in several countries published a mark for every player after a match, based on the reporter's judgement. Italian newspapers' pagelle and the ratings in the French sports daily L'Équipe are long-standing examples.

Those marks were subjective by design. They captured an expert's overall impression but could not be reproduced or checked. As event data became widely available, data providers began generating ratings automatically, applying the same formula to every player in every match. The scale stayed familiar, but the method changed completely.

How an Algorithmic Rating Is Built

Most automated rating systems follow a similar logic, although the exact formulas are proprietary and differ between providers.

  • A starting value: each player begins the match with a baseline score, typically somewhere in the middle of the scale.
  • Positive events: actions such as goals, assists, key passes, successful dribbles, tackles won and interceptions add to the score.
  • Negative events: actions such as losing possession, missed chances, fouls conceded, errors leading to shots or goals, and cards subtract from it.
  • Weighting by context: the same action can be worth more or less depending on where it happened, how dangerous it was and the player's position.
  • Scaling: the final total is fitted to the familiar zero-to-ten range, with most players in most matches landing in a narrow band around the middle.

Some systems also include team-level factors, such as the result or the team's overall performance. Others apply a minimum playing time before a rating is shown, because a player who appears for a few minutes produces too few events for a meaningful score.

Why Position Adjustments Matter

A centre-back and a striker do completely different jobs, so a rating system has to judge them differently. Without adjustment, attackers would dominate every match, because goals and chances carry the largest values.

Rating models therefore weight actions by position. A clearance from a centre-back counts for more than the same action by a winger, and a successful dribble by a forward counts for more than one by a defender. This makes ratings more comparable within a position than across positions, which is an important point when reading them.

The Biases Built Into Ratings

Because ratings are built from recorded events, they inherit the limits of event data. Several systematic biases follow.

Volume bias. Ratings reward involvement. A player who touches the ball often has more chances to accumulate positive actions. Possession-dominant midfielders tend to rate well, while a forward in a low-possession team can do his job well and still receive a modest score.

The invisible defender. Good positioning often prevents events from happening at all. A centre-back who reads the game well may stop attacks before a pass is played, leaving little in the data. A defender on a team under constant pressure, by contrast, records many clearances, blocks and tackles, which can inflate his rating even in a defeat.

Goal weighting. Goals and assists carry large values, so a single goal can lift an otherwise poor performance into a high rating. A player who links play well for 90 minutes without a direct goal contribution may rate lower than a teammate who scored once and did little else.

Goalkeepers on busy days. A goalkeeper who makes many saves usually rates highly, but a high save count often reflects a team that conceded many shots. The rating rewards the workload as much as the quality.

Short appearances. Substitutes who play only a few minutes can receive extreme ratings from one or two events, which is why many systems set a minimum playing time.

What Ratings Do Well

The biases are real, but ratings also have genuine strengths. They are consistent: the same formula applies to every player in every match, so a rating is never influenced by a reporter's mood, reputation or allegiance. They are fast, available as soon as the event data is complete. They also cover far more matches than human reviewers could, including lower divisions and smaller leagues.

Ratings are particularly good at flagging outliers. A player who consistently rates well above his position's average over a season is usually doing something valuable, even if the rating cannot explain exactly what. A player who consistently rates poorly despite a strong reputation is worth a closer look in the underlying data.

They also make trends visible. A sudden drop in a player's average rating across several matches can signal a change in role, a loss of fitness or a tactical shift around him. The rating does not identify the cause, but it tells an analyst where to look. Used as an early-warning signal rather than a final judgement, a rating earns its place in the toolkit.

Ratings Versus Possession Value Models

Analysts increasingly use possession value models, such as expected threat and other action-valuation frameworks, which measure how much each action changes a team's chance of scoring or conceding. These models share some ideas with ratings, since both assign value to individual actions.

The difference is in purpose and transparency. Possession value models are built to estimate the value of actions in a consistent, testable way. Ratings are built to produce a single, readable match score, and their weightings are usually tuned to match human intuition about who played well. A rating is a summary, while a possession value model is closer to a measurement.

A Worked Example

Imagine two hypothetical centre-backs in the same round of matches. Defender A plays for a team that spends most of the match defending deep. He makes many clearances, wins several aerial duels and blocks two shots. His team loses narrowly. Defender B plays for a dominant team that concedes almost no chances. He makes few defensive actions but completes many passes, a handful of them forward through midfield. His team wins comfortably.

A typical rating system may score Defender A higher, because he recorded more positive defensive events. That does not mean he played better. Defender B may have prevented attacks through positioning and contributed more to his team's control. The ratings describe the volume and type of events each player recorded, not the full quality of their performances.

How to Read Ratings Sensibly

Ratings are useful when read with their limits in mind:

  • Compare within a position: a rating is most meaningful against other players in the same role.
  • Use averages over many matches: a season average smooths out the volatility of single games.
  • Check the underlying events: the passes, duels, shots and chances behind a rating explain why it is high or low.
  • Consider the team context: players in dominant teams and players in defensive teams accumulate different kinds of events.
  • Do not mix providers: different systems use different formulas, so a rating from one source cannot be compared directly with a rating from another.

RubiScore's match pages list the underlying events and statistics for each player, which makes it possible to see what a rating is likely to be built from and where it might mislead.

Common Misconceptions

  • "A higher rating means a better performance." It means more, or more valuable, recorded events, which is not always the same thing.
  • "Ratings are objective because they are calculated." The formula is consistent, but the choice of weights reflects the designer's judgements.
  • "A rating can compare a goalkeeper with a striker." Position adjustments make cross-position comparisons unreliable.
  • "One match rating shows a player's level." Single-match ratings are noisy, especially for players with limited minutes.

The Takeaway

Algorithmic player ratings turn event data into a single, familiar score. They are consistent, quick and easy to compare, which explains their popularity. They also inherit the blind spots of event data, rewarding volume and visible actions while missing positioning and off-ball work. Read as a summary of recorded events rather than a final verdict on a performance, a rating can be a useful starting point for deeper analysis.

In Media
Publications
Staff
Al Leejam Newsletter
Endurance Publication


December 2007