[mumo]
A free tool for doing multimodal interaction research.
What it does
- Transcription
- Multimodal transcription for linguistics, conversation analysis, animal behavior, and other fields.
- Conversation Analysis
- Built for CA work: Jeffersonian transcription and overlapping speech.
- Audio Analysis
- Run voice activity detection (VAD) to automatically segment speech. Manually adjusted annotations can snap to audio events.
- Annotation tiers
- Define tiers for speech, gesture, gaze, and other modalities. Compatible with ELAN's tier and constraint model.
- Timeline
- Synchronized waveform and annotation timeline. Smooth rendering even with >10k annotations.
- Extensible
- Define your own transcription conventions. Define structured textual annotation patterns. Write your own plugins.
- Export
- Export to SQLite for analysis, or to EAF for compatibility with existing workflows.
Project status
mumo is under active development.