Skip to main content
Medera Multimodal Sensing is specifically designed for behavioral and mental health. The engines below provide real-time access to clinical signal channels. Selecting the right engine depends on your use case.
Review the Languages page to see language support per surface and the Architecture page for engine deep dives.

Medera Multimodal engines

Vocal Acoustic Engine

F0, jitter, shimmer, HNR, MFCC, prosodic features, and clinical markers from speech.

Facial Physiological Engine

rPPG-derived HR, BP, HRV (SDNN, RMSSD, LF/HF, SD1, SD2), respiration, stress.

Neurobehavioral Construct Computer

15 named neurobehavioral constructs across 5 domains fused from facial, vocal, and assessment context.

Architecture

medera/multimodal.svgthree engines · one construct computerMultimodal sensingvocal acoustics + facial physiology + linguistic contentVocal acoustic engineF0 · jitter · shimmer · HNRSpeaking rate · prosodic flatnessDepression / anxiety / distress indicesFacial physiological engineMediaPipe 468-point landmarksrPPG → HR · HRV (SDNN/RMSSD/LF-HF)BP · RR · stress index · ANS balanceLinguistic content engineTranscript · facts · spansFirst-person ratio · absolutist wordsLexicon density · cognitive distortionNeurobehavioral Construct Computer15 constructs · 5 domains · zero overlapNegative valence · Positive valence · Cognitive systems · Social processes · Arousal & regulatory

Engine functionality


Endpoints

Confidence bands


What’s next

Architecture

Pipeline deep dive from raw signal to canonical clinical payload.

Quickstart

Stream your first multimodal session end-to-end.

Neurobehavioral Construct Constructs

15 named constructs across 5 domains.

Co-Therapy Agent

Agent-level documentation for in-session multimodal.
Contact us if you need help selecting the right engine or have questions about configuring requests. medera.info/contact