Evaluation Protocols for Spatial Meaning
SUMMARY
How to test whether the installation is technically responsive, spatially intelligible, narratively coherent, and socially viable.
DETAIL
Evaluation should separate infrastructure performance from conceptual success. A participant may become confused because tracking is delayed, because the field map is incoherent, because too many layers compete, or because the intended ambiguity is too weakly signaled. These are different failures and require different tests.
Technical measures include position and orientation error, motion-to-audio latency, dropout frequency, source stability, calibration drift, rendering artifacts, and synchronization between shared layers. These tests establish whether movement can plausibly function as a causal interface.
Interaction measures examine whether participants discover the orientation mapping, intentionally approach or avoid fields, recover a previously encountered region, and adapt their movement after hearing a transition. A controlled traversal can compare expected and observed behavior without requiring participants to describe the system in technical language.
Semantic measures test whether spatial adjacency communicates an intended relationship. After traversal, participants can reconstruct a rough map, group neighboring concepts, describe what changed at a boundary, or identify which fields felt related or contradictory. Divergent interpretations may be acceptable, but a transition designed as synthesis should not consistently be perceived as arbitrary noise.
Narrative measures examine continuity, repetition, path dependence, and the effects of drift. Researchers can compare deterministic and generative versions of the same map, or compare current-position-only behavior with trajectory-aware behavior.
Social measures include awareness of other participants, perceived isolation, unwanted influence, privacy expectations, and whether shared moments are recognizable. Accessibility evaluation should include participants with varied hearing, mobility, sensory tolerance, and familiarity with spatial audio.
Useful evidence combines movement logs, renderer logs, observation, map reconstruction, interviews, and task-based comparisons. Immersion is only one outcome. Agency, comprehension, comfort, curiosity, recall, and recoverability should be measured separately.
WHY THIS EXISTS
Supports research studies, prototype comparison, acceptance criteria, debugging, and product validation.
SOURCE CONTEXT POINTERS
- /concepts/position-aware-audio-installation/RESEARCH_DIRECTIONS.txt
- /concepts/position-aware-audio-installation/RISKS_AND_CONTRADICTIONS.txt
- /concepts/position-aware-audio-installation/PATTERNS.txt
EVIDENCE QUESTIONS
- evaluation methods interactive spatial audio installation navigation presence comprehension user study (semantic): The returned corpus mainly described prototype potential rather than established study methods, so this node uses conservative evaluation logic derived from the installation's explicit failure modes