For players · 9 min read
How do chess ratings work? Elo, Glicko and your club rating
Why beating a stronger player is worth more, what the question mark after a new rating means, and why your online number and your club number disagree.
By Castle Gambit · Updated
On this page
On this page
The short answer
A chess rating is a prediction: the bigger the gap between two players’ ratings, the more likely the higher-rated one is to win, and after each game both ratings move by how much the result surprised the system. Elo, which FIDE uses, does this with a single number; Glicko and Glicko-2, used by chess.com and Lichess, also keep track of how sure they are.
Key takeaways
- A rating is a prediction: with a 200-point gap, the stronger player is expected to score about 76%.
- Beat someone rated above you and you gain more; lose to someone rated below you and you drop more.
- Glicko and Glicko-2 also track how uncertain a rating is, so new ratings move fast and settled ones move slowly.
- Ratings from different places (FIDE, US Chess, chess.com, Lichess) measure different pools of players, so they don’t convert.
- Castle Gambit’s social rating uses Glicko-2 on casual games played in person, and stays hidden for your first five games.
Every chess player eventually gets a number, and then spends years quietly obsessing over it. This guide explains where the number comes from, why it moves the way it does, and why the one on your phone and the one at your club will never agree. There’s a little arithmetic. It’s the friendly kind.
What a rating actually is
A chess rating isn’t a measure of talent, and it isn’t a score you add to. It’s a prediction: given two players’ ratings, how is a game between them likely to go? Every rating system in chess does the same three things:
Predict. Before a game, the ratings say how much each player is expected to score.
Compare. After the game, the system compares what happened with what it predicted.
Adjust. Both ratings move by how surprised the system was. No surprise, small change. Big upset, big change.
That’s why the number on its own means less than you’d think. What matters is the difference between two ratings, and it only means something between players in the same pool: people who play each other, or play people who play each other.
Elo: the expected score
The system most of chess grew up with is named after Arpad Elo, a physics professor and chess player whose method was adopted by the US Chess Federation in 1960 and by FIDE, the international chess federation, in 1970. Its central idea is the expected score: what a player should score against an opponent, where a win counts 1, a draw ½ and a loss 0.
The expected score depends only on the rating gap. Equal ratings: 50% each. Rated 200 points higher: about 76%, or roughly three points out of four. Rated 400 higher: about 91%. The curve is steep in the middle and flat at the ends, which is why nobody is ever guaranteed a win, however big the gap.
For the curious, here’s the formula, for a player rated R against an opponent rated Ropp:
E = 1 / (1 + 10^((R_opp − R) / 400))
The 400 sets the scale: it’s why ratings run in the hundreds and thousands rather than single digits.
How a rating goes up and down
After a game, Elo moves your rating by the difference between what you scored and what you were expected to score, multiplied by a number called K:
new rating = old rating + K × (score − expected score)
Say you’re rated 1500 and you play someone rated 1700. You’re expected to score about 0.24. With a K of 20:
| Result | Your score | Change | Why |
|---|---|---|---|
| Win | 1 | +15 | A big surprise: 20 × (1 − 0.24) |
| Draw | ½ | +5 | Better than expected: 20 × (0.5 − 0.24) |
| Loss | 0 | −5 | Roughly what was predicted: 20 × (0 − 0.24) |
Beating a stronger player is worth a lot; losing to one costs a little. Losing to a weaker player costs a lot. And whatever one player gains, the other loses, when both use the same K.
K sets how fast ratings move. FIDE uses 40 for newcomers and most juniors, 20 for most players, and 10 for players who have reached 2400. A bigger K reacts faster; a smaller one is steadier.
Glicko and Glicko-2: how sure are we?
Elo has a blind spot: it’s equally sure about everyone. A player with three games gets the same treatment as a player with three thousand. In the 1990s, the statistician Mark Glickman fixed that with the Glicko system, and a few years later with Glicko-2.
Glicko gives every player a second number, the rating deviation (RD): how uncertain the system is about your rating. Your real strength is very probably within about two RDs of your rating either way. A brand-new player might be 1500 with an RD of 350, meaning “somewhere between about 800 and 2200. We have no idea yet.”
- RD shrinks as you play. Every game tells the system more, so the range narrows.
- RD grows when you don’t. Take a long break and the system becomes less sure of you again.
- Uncertain ratings move fast; settled ones move slowly. Your first games swing your rating around, and that’s the system learning, not you getting worse.
- Uncertain opponents count for less. Beating someone the system knows little about tells it less than beating someone it knows well.
Glicko-2 adds a third number, volatility: how erratic your results are. A player who is improving fast, or having a very odd month, gets a rating that’s allowed to move more.
Provisional ratings and the question mark
Because early ratings are guesses, most systems mark them as provisional until there are enough games behind them. Lichess shows a question mark next to a rating it isn’t sure of yet. US Chess marks ratings based on only a few games as provisional. FIDE doesn’t publish a first rating at all until you’ve played enough rated games against rated opponents.
Castle Gambit’s social rating goes a step further: for your first five games you don’t see a number at all, only tally marks. After the fifth you get a number with a question mark (like “1489?”), and the question mark goes once the system is sure enough.
What you see
5 more games to your first rating.
What the rating system knows
Notice how much the number jumps early on (1500, then 1302, then 1501) while the range is wide, and how calm it gets by the end. That’s exactly why it’s hidden at first: an early number mostly measures luck.
Who uses which system
The big rating pools, and the systems behind them:
| Who | System | What it rates |
|---|---|---|
| FIDE | Elo | Rated over-the-board games worldwide, with separate standard, rapid and blitz lists |
| US Chess | Its own Elo-based system | Rated games in the US, with separate ratings for faster time controls and for online play |
| chess.com | Glicko | Online games, separately for each time control |
| Lichess | Glicko-2 | Online games, separately for each time control |
| Castle Gambit | Glicko-2 | Casual games played in person at Castles (chess clubs) |
The details inside each (starting numbers, floors, K-factors, how they handle new players) differ, and they change from time to time. The idea underneath is always the same: predict, compare, adjust.
Why your online and over-the-board numbers differ
It’s the most common question at any club: “I’m 1600 online, what am I over the board?” The honest answer is that there’s no conversion, because a rating only means something inside its own pool, and the pools are different in every way that matters:
- Different people. Who plays online and who plays at a club overlap, but they aren’t the same crowd.
- Different time controls. Most online games are much faster than club games, and blitz skill and slow-chess skill aren’t the same thing.
- Different starting points and rules. Each site and federation starts new players somewhere different and moves ratings at a different speed.
- A different game, a little. A real board, a real clock and a real person across the table change how people play.
So your online rating is useful as a rough starting point and nothing more. That’s how Castle Gambit uses it: link a chess.com or Lichess account and your social rating starts near it, with plenty of uncertainty, so your first real games correct it quickly.
What’s a good rating?
It depends entirely on the pool, which is an unsatisfying answer and the true one. The average rating on any list depends on who’s on it and how that system starts people off, so a number that’s strong in one pool can be ordinary in another.
The top end does have landmarks. FIDE’s titles need a rating to go with them: FIDE Master needs 2300, International Master 2400, and Grandmaster 2500, the last two along with strong results called norms.
For the rest of us, the useful comparison isn’t anybody else’s number. It’s yours, a few months ago. A rating that’s slowly climbing means you’re getting better, which was the point all along.
09Castle Gambit’s social rating
Castle Gambit keeps its own rating (we call it a social Elo; the app just says “rating”) for the games most club players play most of the time: casual games across a real board. It’s separate from FIDE, US Chess, chess.com and Lichess, and it measures one thing: how you play.