neuphonic/neutts-air
Text-to-Speech • 0.7B • Updated • 8.8k • 878
State-of-the-art target speech extractor
Extreme Super-Resolution via Scale Autoregression
Generate large-scale 3D models with spatial sparse attention
Train neural networks live and understand how they learn
Voice Activity Detection using MarbleNet model
Filter multilingual data for high-quality language models
Transcribe speech and highlight emphasized words