Towards Quantifying Benchmark Optimization in ASR Models Paper • 2608.19936 • Published 7 days ago • 11
view post Post 30230 Excited to share that I've joined the Hugging Face Fellows program! 🤗Looking forward to contributing to & working more closely with the open-source ecosystem - huge thanks to everyone who's supported me on this journey! 🚀 See translation 🤗 10 10 🚀 2 2 + Reply
Towards Robust and Generalizable Lensless Imaging with Modular Learned Reconstruction Paper • 2502.01102 • Published Feb 3, 2025
LenslessMic: Audio Encryption and Authentication via Lensless Computational Imaging Paper • 2509.16418 • Published Sep 19, 2025 • 1
Treble10: A high-quality dataset for far-field speech recognition, dereverberation, and enhancement Paper • 2510.23141 • Published Oct 27, 2025 • 5
Open ASR Leaderboard: Towards Reproducible and Transparent Multilingual and Long-Form Speech Recognition Evaluation Paper • 2510.06961 • Published Oct 8, 2025 • 14
view post Post 6515 Trained a model for emotion-controllable TTS based on MiMo audio on LAION's dataset.Still very early and does have an issue with hallucinating but results seem pretty good so far, given that it is very early into the training run.Will probably kick off a new run later with some settings tweaked.Put up a demo here: https://huggingface.co/spaces/mrfakename/EmoAct-MiMo(Turn 🔊 on to hear audio samples) See translation 5 replies · 🔥 12 12 + Reply