LAION eV
AI & ML interests
open multi-modal foundation models and datasets for their creation; scaling laws, model evaluation; fully local, sovereign model deployment, personalized assistants and open local agentic systems
Recent Activity
Vocal Burst Segments — Samples
Explore labeled vocal burst audio samples
Burst LoRAs trained on Gemini labels vs the shipped ones
Compare audio burst adapters and view detection results
Gemini Burst Verification
Explore annotated vocal burst samples with audio playback
Gemini as a vocal-burst annotator
Explore Gemini’s burst annotations on drama audio clips
DramaBox — can it produce screams and shrieks?
Generate custom scream or shriek audio from scene text
Burst Guidance Listening
Scream and Shriek — what the burst adapters actually produce
Explore generated scream & shriek audio at varying strengths
MOSS Vocal Burst Recipes
Listen to vocal burst samples and see detection results
Synthesised Vocal Bursts — does it work?
Listen to synthetic vocal bursts and assess their labels
Emotion Continuity — 20 Transitions
Rate emotional crossfades in AI‑generated speech clips
MOSS Quality Adapter Listening Test
Rate and compare synthetic speech with different quality adapters
Crossfade v2 — Listening Study
Compare emotional voice transitions and rate smoothness
Emotional Transitions — Listening Study
Rate emotional voice transitions by listening to audio clips
MOSS Voice-Acting — Emotion LoRA Merge-Weight Ablation
Adjust emotion intensity in synthetic speech with LoRA weights
MOSS Voice-Acting — Turning the Emotion Up at Inference Time
Adjust emotional tone in generated speech
What Makes a Generated Voice Sound Emotional
A controlled 4-factor study on the MOSS voice-acting model
MOSS Voice-Acting — Technical Report
Generate expressive speech from stage directions
MOSS Voice-Acting — Emotion LoRAs vs Baseline
Listen to emotion‑controlled speech samples
MOSS Voice-Acting v2 — Round-3 Evaluation
Explore and compare voice‑acting model performance with audio samples
MOSS Voice-Acting v2 — Round-2 SFT Evaluation
Generate emotional voice clips with precise timing
MOSS Voice-Acting v2 — Round-2 SFT Evaluation
Generate emotion‑controlled speech from timed scripts
MOSS Voice-Acting SFT — Validation Samples
Listen to and compare voice‑acting model samples
CoCa
Generate descriptive captions for images using advanced AI technology