neuphonic/neutts-air
Text-to-Speech • 0.7B • Updated • 6.3k • 879
State-of-the-art target speech extractor
Extreme Super-Resolution via Scale Autoregression
Generate large-scale 3D models with spatial sparse attention
Native mixed-anchor matrix-completion audit
Voice Activity Detection using MarbleNet model
Filter multilingual data for high-quality language models
Transcribe speech and highlight emphasized words