Evaluation asks whether interpretation works for actual visitors in its real context—and makes that evidence useful for decisions. It belongs before, during, immediately after, and well after development.
The evaluation cycle: investigate prior understanding, test developing ideas, correct the installed experience, and assess outcomes after launch. Do not wait until every expensive decision is irreversible.
1. Begin with intended outcomes
Define what visitors should be able to notice, understand, question, navigate, access, discuss, or choose. Add operational and ethical outcomes: staff can support the experience, corrections are manageable, and sensitive interpretation is received with appropriate context.
Each outcome needs an evaluation question. “Visitors engage deeply” is too vague. “Can visitors identify the visual evidence behind the interpretation?” can be researched.
2. Front-end evaluation
Front-end work happens before a solution is defined. It explores prior knowledge, language, associations, expectations, misconceptions, interests, barriers, and the relevance of possible themes. Methods include short interviews, concept sorting, observation, community workshops, and review of existing evidence.
3. Formative evaluation
Formative evaluation tests developing components while they can still change. Use rough labels, storyboards, audio drafts, route prototypes, clickable interfaces, and temporary signage.
- Can visitors find and start it?
- Do they understand the main idea?
- Does audio direct attention to the intended evidence?
- Can different visitors control and navigate it?
- Does the tone suit the subject and place?
- What happens in groups?
4. Remedial evaluation
After installation, inspect problems created by the actual environment: glare, sound, congestion, blocked sightlines, weak connectivity, misplaced QR codes, staff workflows, and objects that changed position. Fixing these quickly can have more impact than adding new content.
5. Summative evaluation
Summative work assesses the implemented experience against its intended outcomes. Combine behavior, visitor accounts, comprehension, access, emotional response, institutional goals, and operations. Use it for accountability and learning—not merely to produce a positive headline.
6. Choose methods that fit the question
| Question | Useful methods | Limit |
|---|---|---|
| Can visitors find and use it? | Observation, task testing, support logs, funnel data | Does not explain meaning by itself |
| What did visitors understand? | Open interview, recall, explanation in their own words | Question wording can lead answers |
| How did it shape looking? | Observation, visitor-led walkthrough, interview beside object | Research presence may affect behavior |
| Who is excluded? | Accessibility review and testing with relevant users | A small sample does not represent every need |
| What happens at scale? | Aggregated usage data, survey, operational records | Patterns do not automatically explain causes |
7. Interpret evidence carefully
Separate finding, inference, and recommendation. Look for convergence across methods. Preserve important minority or access findings even when they are not common. Do not claim causation from an uncontrolled before-and-after comparison.
Document who participated and who did not. Evaluation is evidence for judgment, not a machine that removes institutional responsibility.
8. Where GuideSofia fits
GuideSofia supports prototyping, controlled pilots, route and content variants, multilingual delivery, and privacy-conscious usage signals. Those signals become more useful when interpreted alongside research rather than treated as a complete account of the visit.
9. Evaluation template
- Decision and intended outcome.
- Evaluation question and success evidence.
- Audience and recruitment approach.
- Method, context, consent, and data handling.
- Finding, confidence, and limitations.
- Interpretation and alternative explanations.
- Recommended action, owner, deadline, and retest.
Frequently asked questions
What is formative museum evaluation?
Formative evaluation tests developing ideas and prototypes early enough to change them. It can examine comprehension, navigation, usability, relevance, accessibility, and emotional fit.
What is summative museum evaluation?
Summative evaluation studies the implemented experience and the extent to which it achieved intended outcomes. It can inform accountability, future projects, and later improvements.
How many visitors are needed for evaluation?
It depends on the method and decision. Small qualitative rounds can reveal major usability and comprehension problems; claims about population prevalence require appropriate sampling and statistical design.
Is visitor satisfaction enough?
No. Satisfaction is useful but does not show whether visitors noticed intended evidence, understood ideas, navigated successfully, felt represented, or could access the experience.