Beyond the Score: What Football Analytics Teaches Assessment

Share via:

Why looking past pass/fail matters as much as looking past the full-time score.

With the 2026 World Cup now behind us, most of what we remember will be the scorelines.

Spain beat Argentina 1- 0 

England finished third after beating France 6 – 4 

Scotland… actually, let’s move on

Those are the results that will appear in the history books.

But in modern football the final score is actually only part of the story. 

When the Final Whistle Blows – The Analysis begins 

Nowadays, most football teams employ teams of data analysts, who gather and analyse an extraordinary range of data: possession, passes completed, chances created, expected goals, even where on the pitch the play actually happened.

They gather a huge dataset not just at a team level, but at an individual player level. Some of those numbers can feel a bit baffling, if you’ve ever nodded along to a conversation about xG while secretly wondering what on earth it really means, you’re not alone.

It can feel completely alien to the beautiful game that most of us fell in love with, a game of passion and flair and enjoyment, 

But for coaches and analysts, these extra layers of data really matter in a sporting world where success is increasingly defined by the thinnest of margins. 

The result tells us what happened, when the final whistle blew, and who the winner was. 

The deeper data helps them understand how it happened, what went well, what did not go as planned. It begins to form the foundations for how they will approach the next game. 

 The Assessment Scoreline 

In assessment, we have our own version of the scoreline.

Candidates pass or fail. 

Cohorts hit certain thresholds. 

We produce overall statistics and trend lines. 

Just like the final score in the football game, these things matter, these are the numbers that are written on our certificates, and our stakeholder reports. 

They tell us who met the standard, who didn’t, and how cohorts compare over time.

But, just like in football, they only give us part of the story.

Behind every assessment result sits a huge amount of additional information. 

  • How candidates performed on individual items
  • How long they spent on different questions or sections
  • Where they changed answers or abandoned attempts
  • Which accessibility or support tools they used
  • How different centres, regions or cohorts behaved

If we only ever look at the final scores, we don’t just miss that bigger picture. We risk over-simplifying a complex learning and assessment journey.

Beyond Pass/Fail

When we look beyond the final result, we can ask more useful questions, in much the same way that the football analysts look at possession, xG and player heat maps. 

For example: 

  • Item performance: Which questions are consistently causing confusion, or failing to differentiate between levels of ability?
  • Timing patterns: Are candidates spending longer on certain sections than we expect, suggesting unnecessary cognitive load?
  • Completion behaviour: Are there groups who regularly submit late, or fail to complete assessments at all?
  • Accessibility usage: Are support tools genuinely helping the candidates who rely on them?
  • Cohort and regional differences: Which cohorts or centres are consistently ahead or behind the curve?

These are the assessment equivalents of asking: 

  • Did we control the midfield? 
  • Were our full backs too exposed? 
  • Did our pressing work?

They don’t replace the result. 

They help to explain it.

They help the managers to learn from the past and build for the future. 

A Game of Two Halves

The most important thing about these deeper layers of assessment data is not the numbers themselves, but what they enable us to do next. 

When we understand more than just the scoreline, we can improve questions that aren’t working as intended. We can identify where learners are struggling across the course, not just in the final exam. We can check whether our support arrangements are genuinely helping the people who need them and adjust the balance and timing of assessments.

And we can see risks emerge early in the process rather than waiting for the point when they stop being risks and become problems. 

In defence of the grumpy fan

As a supporter, this new world has turned a lot of post-match interviews and pundits’ commentary into nothing more than a barrage of stats, often used to try to hide the fact that, on the day, the team just was not good enough. 

I don’t really care that my team’s possession stats were much better than the opposition’s, I care that we lost 2-0. 

But behind the scenes, this information is incredibly powerful.

It changes the conversation from: “We lost 2-0. The players weren’t good enough.”

To: “We lost 2-0. We created far fewer chances than usual despite dominating possession. Why?”

And once we start asking why, improvement becomes possible.

The same applies in assessments. Instead of saying “more people failed than we would expect”, we can look into the data to find patterns that will not only tell us why more people failed, but what we need to do differently next time.

Results will always matter most, on the pitch and in assessment. But the final score only tells us what happened.

The data tells us why.

And if we want a better result next time, “why” is usually the more important question.

This is exactly the thinking behind Cirrus Data & Insights

Share via:
Topics
Picture of Robert Burns
Robert Burns
Robbie works with assessment organisations on the practical realities of delivering high-stakes assessments, with a particular focus on operational resilience, assessment integrity and candidate experience in digital delivery models.
Would you like to receive Cirrus news directly in your inbox?
More posts in Better Assessments
Better Assessments

Choosing the Right Proctoring Model for Your Assessment

Everyone wants to know the best way to proctor an exam. There isn’t one. Every qualification carries its own purpose, stakes, candidates and constraints, so the useful question is not which model is best, but which model fits, and whether you can explain why you chose it.

Read More »
Better Assessments

How Secure Is Secure Enough?

There is no universal definition of secure enough. The answer depends on what the task exposes, what is at stake if the result is wrong, and the scale you are working at. The third article in The Proctoring Question sets out how to match the control to the risk, and what it costs when you get it wrong.

Read More »

Know exactly where your assessment operation stands

Your free personalised report in 4 minutes

Answer 12 questions across strategy, delivery, design and data and get a clear, personalised breakdown of where your assessment operation is strong and where to focus next.