Calculate Strength of Schedule in Soccer: 5 Steps and a Live BetsyScore Test

Strength of schedule (SOS) measures how difficult a team’s opponents are, based on those opponents’ own quality and results. We use it because a won-loss record alone does not tell the full story: SOS adjusts how we read a record and feeds directly into seeding decisions, power rankings, and win-probability models. Several calculation methods exist, and the right one depends on what you are trying to predict.
TL;DR:
- A schedule that faces many strong opponents increases a team’s real strength, which is captured by the z-score of its SOS, not just win-loss record.
- Different models like NCAA’s Power Index, Elo ratings, and fixture difficulty tools suit various prediction and ranking purposes, each with unique strengths and limitations.
- SOS calculations rely on opponent ratings and their opponents’ ratings, weighted more heavily for direct opponents, and can be affected by small sample sizes and venue effects.
- Combining SOS with live data such as expected goals, recent form, and momentum enhances prediction accuracy for both rankings and match outcomes.
- Use a z-score threshold of ±1 to identify when schedule difficulty meaningfully impacts team evaluation, but always consider data stability and independence for reliable conclusions.
Table of Contents
- What strength of schedule measures and where it matters
- How to calculate strength of schedule step by step
- Comparing NPI, Elo ratings, and fixture difficulty models
- Reading and applying SOS in rankings and predictions
- Limitations and common pitfalls to avoid
- Putting SOS to work with live match data
- A checklist for trusting an SOS figure
- Try BetsyScore for live predictions that combine SOS with live metrics
- FAQ
- Sources
What strength of schedule measures and where it matters
A team that wins eight of ten matches against weak opposition has not proven the same thing as a team with the same record against strong opposition. Raw results flatten that difference, which is exactly why SOS exists: it puts a number on the quality of the path a team walked to get its record.
We see SOS used in three recurring contexts:
- Power rankings: analysts adjust a team’s position based on who it has actually played, not just how many points it has collected.
- Selection and seeding committees: panels weigh schedule strength alongside wins and losses when deciding who qualifies or how they are seeded, a practice the NCAA Power Index documentation formalizes for committee use.
- Predictive models: forecasting systems use SOS as an input feature, since a team’s underlying quality is easier to estimate once opponent strength is accounted for.
A mid-table side that has faced five of the league’s top six teams in its first ten games looks very different once that context is added, even if its points total matches a team that avoided all of them.
How to calculate strength of schedule step by step
Most SOS formulas start with two building blocks: the average rating of a team’s opponents (OR, for “opponents’ rating”) and the average rating of those opponents’ own opponents (OOR, for “opponents’ opponents’ rating”). OOR exists so that beating a team with a weak schedule of its own counts for less than beating a team that has also faced strong competition.
A common weighting treats direct opponents as more informative than indirect ones:
- Build opponent ratings. Use a season-level metric such as goal difference per 90 minutes or non-penalty expected goals per 90 as the baseline rating for every team in the league.
- Average OR across remaining or played fixtures, adjusting for home and away split if the data supports it, since away form is typically weaker across most leagues.
- Average OOR, the opponents of each of those opponents, using the same baseline rating.
- Combine the two with a standard weighting: SOS = 2/3 × OR + 1/3 × OOR.
- Standardize the result into a z-score against the league mean and standard deviation, so a value of 0 means average difficulty and +1 means one standard deviation harder than average.
Fixture difficulty ratings used in fantasy football score each individual game from 1 to 5 using recent form and home or away splits, reviewed on a weekly basis. That per-game view complements the season-level SOS figure described above.
Say a team’s next three opponents carry ratings of 1.4, 0.8, and 1.1 (OR average of 1.1), and those opponents’ own opponents average 0.9 (OOR). The raw SOS is 2/3 × 1.1 + 1/3 × 0.9, which works out to 1.03. Converted to a z-score against a league where the mean SOS is 0.9 and the standard deviation is 0.2, that team’s schedule sits at roughly +0.65, moderately tougher than average.

Comparing NPI, Elo ratings, and fixture difficulty models
Three approaches dominate how SOS gets built into usable tools, and each suits a different job.
- NPI/OPI style indices let a committee tune the weighting between winning percentage and schedule strength, which is why selection bodies favor them: the NCAA Power Index documentation describes configurations such as an 80% weighting on strength of schedule against a 20% weighting on raw winning value.
- Elo-based ratings update after every match rather than at fixed intervals, and the gap between two teams’ Elo ratings correlates strongly with match outcomes, which is why they anchor most probabilistic forecasting systems, as the Elo introduction for football forecasting lays out.
- Fixture Difficulty Rating (FDR) tools score single games rather than a whole season, which makes them well suited to short-run forecasting and fantasy lineup decisions rather than committee-level resume building.
Recent research combining sufficient-dimension reduction of lagged Elo rating histories with a Poisson goals model outperformed simpler baselines when forecasting the 2026 World Cup. Predicting the 2026 FIFA World Cup with sufficient dimension reduction of Elo rating histories
That hybrid approach, pairing an Elo-style measure of opponent strength with a Poisson scoring model, is roughly how modern prediction pipelines fold SOS into a probability rather than treating it as a standalone label.
Reading and applying SOS in rankings and predictions
Selection committees rarely use SOS as a single deciding factor. The NCAA’s description of NET rankings treats schedule strength as one of several resume criteria alongside bad losses, common opponents, and road record, which means two teams with similar win totals can be viewed very differently once their schedules are compared.
In predictive pipelines, SOS typically functions as an input feature sitting alongside Elo and expected goals data, not as a replacement for either. Analysts use it in a few recurring ways:
- As a tie-breaker when two teams otherwise look equal on points or win percentage.
- As a weighting adjustment inside a power ranking formula, lowering the implied quality of a record built against weak competition.
- As a sanity check on live win probabilities, flagging when a short unbeaten run was built against a soft schedule.
Pro Tip: Treat a z-score of ±1 or higher as the threshold where SOS is likely to change your read on a team; smaller differences are often noise.
Limitations and common pitfalls to avoid
SOS is only as reliable as the data behind it, and a few recurring problems undercut it early in a season or when applied carelessly.
- Small-sample bias: early-season ratings swing heavily after just one or two results, which makes SOS unstable until a league reaches roughly a third of its schedule.
- Circularity: if opponent ratings are derived from the same results SOS is meant to explain, the measure can reinforce its own errors rather than correct them.
- Venue effects: home and away splits are not symmetric across leagues, so a raw average that ignores venue can overstate or understate true difficulty.
- Isolation from context: SOS says nothing about injuries, fatigue, or tactical matchups, so it works best alongside expected goals and recent form, not in place of them. Our soccer match analysis checklist for coaches covers which of those inputs matter most.
Putting SOS to work with live match data
Season-level SOS tells you about the past schedule; live inputs tell you what is happening right now. Our platform sits between those two views, combining expected goals, recent form, and head-to-head records into win-probability percentages, alongside a minute-by-minute momentum read that shows which side is controlling a match as it unfolds.
A practical workflow looks like this:
- Compute opponent ratings and standardize them into a z-score for the team you are tracking.
- Feed that SOS figure alongside Elo-style ratings into a probability estimate, following the approach our prediction algorithm guide walks through.
- Compare your pre-match expectation against live momentum on our live scores page to see whether the match is tracking your schedule-adjusted hypothesis.
A checklist for trusting an SOS figure
Before leaning on any SOS number, check three things: is the sample large enough to be stable, are the opponent ratings calculated independently of the record you are judging, and does the figure account for home and away splits. If any of those fail, treat the number as a rough signal rather than a verdict.
Fans should trust SOS most when it confirms what expected goals and form already suggest, and question it most when it stands alone against those other signals. We would rather see more analysts publish their weighting choices and opponent-rating methods openly, since reproducible SOS work is far more useful than a single unexplained number.
— Aria
Try BetsyScore for live predictions that combine SOS with live metrics
Once you have a schedule-adjusted view of a team, the next step is testing it against something live. The platform pairs real-time scores with AI-generated win probabilities built from expected goals, form, and head-to-head records, updating every few seconds so you can see whether a tough schedule is actually showing up in results.
Two ways to put this to use: check a team’s upcoming run of opponents, then compare your own SOS read against our win-probability figures on the AI predictions page, or watch a live match unfold on our live scores page to see whether in-game momentum lines up with the schedule context you calculated.
FAQ
Where do you put your weakest players in soccer?
This is a tactical lineup question rather than a schedule-strength one: coaches generally position weaker individual players where the system can shield them, often in deeper or wide roles with more defensive cover. It has no direct connection to SOS, which measures opponent quality across a whole schedule rather than player placement within a single match.
What are some soccer strengths?
In a tactical sense, team strengths typically include passing accuracy, pressing organization, set-piece delivery, and defensive structure. In the schedule-strength sense this article covers, a team’s “strength” refers to its opponents’ ratings, not its own playing qualities.
Which soccer league is strongest?
League strength is usually assessed through aggregated club ratings such as Elo, which update after every match and correlate strongly with results, as described in the Elo introduction for football forecasting. There is no single official strength of schedule ranking across all leagues, since methods and data sources vary by competition.
What are the best strength exercises for soccer players?
Physical strength training for soccer players typically centers on compound lower-body movements, such as squats and deadlifts, alongside single-leg and core stability work to support sprinting and change of direction. This is a conditioning topic separate from strength of schedule, which is a statistical measure of opponent difficulty rather than a fitness metric.
How is strength of schedule different from fixture difficulty rating?
Strength of schedule is typically a season-level or multi-game average of opponent quality, while a fixture difficulty rating scores each individual match on a simple scale, such as the 1 to 5 system used in fantasy football tools, reviewed weekly using recent form and home or away splits. Both rely on similar underlying data but serve different time horizons.
Sources
- NCAA Power Index documentation
- Elo introduction for football forecasting
- Predicting the 2026 FIFA World Cup with sufficient dimension reduction of Elo rating histories
