Every Assessment Is a Set of Trade-Offs 

Share via:

Many of the assessment design conversations that I am involved in at the moment are framed as optimisation problems. How do we improve security without harming experience? And how do we increase accessibility without compromising validity?

These are valid questions, but potentially a little misleading. 

In reality, these are not optimisation problems, they are trade-offs, and not necessarily ones that are neatly resolved.

Every assessment is ultimately a series of decisions about what we choose to prioritise, and what consequences we are willing to accept as a result. The reality is that most of these trade-offs are already being made. The question is whether we are making them consciously.

Assessment systems today are expected to do more than ever. They must be secure enough to withstand scrutiny, accessible enough to be inclusive, intuitive enough to be fair, and robust enough to measure something meaningful.

Individually, each expectation is entirely reasonable. Taken together, however, they can pull assessment design in different directions.

And if we do not address these tensions, they don’t simply disappear, they reveal themselves in the candidate’s experience, in appeals or contested results, and they can lead to a wider mistrust of our assessment processes. 

Whose definition of “normal”are we designing for?

Many of us have had to develop high security environments to detect suspicious behaviour. But the definition of “suspicious” can be narrower than we think.

For some candidates, particularly those with disabilities or accessibility challenges, behaviour that falls outside the norm is not an anomaly, it is a necessity. Assistive technologies, involuntary movements, or the need to step away from the screen can all trigger flags in tightly controlled environments.

At that point, the system isn’t just protecting integrity, it is making a judgement about what “normal” looks like. And when viewed from that perspective, it does pose the question: Whose definition of “normal” has been embedded into the assessment process?

The cumulative weight of assessment security

Security controls are rarely experienced individually. Candidates experience them cumulatively.

On its own, browser lockdown may be fine. Identity verification alone may be fine. Restricted navigation alone may be fine. But together they create a different assessment experience.

Each intervention is defensible. But combined, they can fundamentally change the assessment experience. 

Under pressure, every bit of friction matters.It increases cognitive load, heightens anxiety, and can shift focus away from the task itself.

So, as part of our design process, we need to ask ourselves, at what point does securing the assessment begin to interfere with the very performance we are trying to measure?

When accessibility risks changing the construct itself

Accessibility is often framed as an unambiguous good. 

That is true. 

But in practice, it can also force assessors to make difficult design choices.

Adapting content for accessibility may involve simplifying or removing elements that are not easily translated across formats. In some cases, those elements are peripheral, but in others, they can be central to the skill being assessed.

If, for example, a task is designed to measure the interpretation of complex visual information, what happens when that visual complexity is removed?

We may improve access, but are we also changing what is being measured? And if we are, can we still claim that the assessment is measuring the same thing?

Authentic assessments create their own barriers

I have, for many years, been a passionate advocate for more “authentic” assessments, assessments which more accurately mirror the real world. 

Whilst this is the way forward for many organisations as we seek to adapt to the digital environment that we now operate in, such assessments can also risk introducing their own barriers. 

It is critical that we design these assessments carefully to ensure that we are not asking candidates to navigate unfamiliar interfaces, manage multiple inputs, and interpret system behaviour alongside the task itself.

While we have a responsibility to prepare candidates for the environments in which they will ultimately work, we also have a responsibility to ensure that we are measuring capability rather than familiarity with a particular interface, level of digital confidence, or tolerance for friction.

Standardisation was never the whole story

For decades, standardisation has been the foundation of fair assessment: the same test, under the same conditions, for everyone.

But accessibility challenges that model. Meaningful inclusion often requires flexibility: adjustments to timing, format or interaction. And the more flexibility we introduce, the harder it becomes to demonstrate that outcomes remain comparable.

It is here that we can often put strain on our processes and on our teams that support them. Not because flexibility is wrong, but because many of our traditional assumptions about fairness were built around uniformity rather than equity.

As if these tensions were not already complex enough, AI is adding another layer to the discussion. Measures designed to prevent inappropriate AI use may increase surveillance, restrict candidate behaviour or alter assessment design. At the same time, failing to address AI may undermine confidence in assessment outcomes. 

Again, there is no perfect solution, only a series of trade-offs that need to be made consciously and transparently

Where do we go from here?

Nothing here is new. 

And none of it means we should abandon security, accessibility, usability, or validity. Quite the opposite. The challenge is recognising that these objectives do not always align perfectly.

There will be times when strengthening security introduces friction. Times when improving accessibility raises questions about comparability. Times when increasing authenticity creates new barriers for some candidates.

The challenge is not to eliminate these tensions, but to acknowledge them, examine them, and make conscious decisions about how they are managed.

Because every assessment reflects a set of priorities.

Trade-offs are inevitable. We simply have to ask ourselves whether we have been deliberate enough in making them.

Hear more at the International e-Assessment Conference 2026

These questions are worth exploring in the open, and that is exactly what the Accessibility Special Interest Group panel is designed to do.

Robbie will be joining fellow assessment professionals on the Main Stage on 9 June for The Future of Assessments: Why Integrity will be Pivotal to Fairness and Inclusion, a discussion on how integrity can support, rather than hinder, inclusive assessment design.

The session runs 11:30–12:00 GMT. Join the conversation. 

Share via:
Topics
Picture of Robert Burns
Robert Burns
Robbie works with assessment organisations on the practical realities of delivering high-stakes assessments, with a particular focus on operational resilience, assessment integrity and candidate experience in digital delivery models.
Would you like to receive Cirrus news directly in your inbox?
More posts in Better Assessments
Better Assessments

Choosing the Right Proctoring Model for Your Assessment

Everyone wants to know the best way to proctor an exam. There isn’t one. Every qualification carries its own purpose, stakes, candidates and constraints, so the useful question is not which model is best, but which model fits, and whether you can explain why you chose it.

Read More »
Better Assessments

How Secure Is Secure Enough?

There is no universal definition of secure enough. The answer depends on what the task exposes, what is at stake if the result is wrong, and the scale you are working at. The third article in The Proctoring Question sets out how to match the control to the risk, and what it costs when you get it wrong.

Read More »

Know exactly where your assessment operation stands

Your free personalised report in 4 minutes

Answer 12 questions across strategy, delivery, design and data and get a clear, personalised breakdown of where your assessment operation is strong and where to focus next.