Whisper Base & Small Β· 60-Clip Listening Gallery
Self-contained Base/Small S10 predictions on 60 audio clips
open multi-modal foundation models and datasets for their creation; scaling laws, model evaluation; fully local, sovereign model deployment, personalized assistants and open local agentic systems
Self-contained Base/Small S10 predictions on 60 audio clips
Humaneness Ears Base/Medium audio listening atlas
Explore and compare voice acting model performance across stages
Whisper Tiny/Base audio scores and burst timing
Explore performance data and audio samples of Humaneness Voice Small
Generate onβdemand scream and shriek voice clips
Generate expressive speech from scripted directions
Explore labeled vocal burst audio samples
Compare audio burst adapters and view detection results
Explore annotated vocal burst samples with audio playback
Explore Geminiβs burst annotations on drama audio clips
Explore generated scream &β―shriek audio at varying strengths
Listen to vocal burst samples and see detection results
Listen to synthetic vocal bursts and assess their labels
Rate emotional crossfades in AIβgenerated speech clips
Rate and compare synthetic speech with different quality adapters
Compare emotional voice transitions and rate smoothness
Rate emotional voice transitions by listening to audio clips
Adjust emotion intensity in synthetic speech with LoRA weights
Adjust emotional tone in generated speech
A controlled 4-factor study on the MOSS voice-acting model