Peer-oriented research on room acoustics, reverberant speech, and robust ASR — each paper paired with an open dataset, benchmark, or tool for full reproducibility.
RIR-Mega: A Large-Scale Simulated Room Impulse Response Dataset for Machine Learning and Room Acoustics Modeling
A large collection of simulated RIRs with a compact, machine-friendly metadata schema, a Hugging Face Datasets loader, validation and checksum tooling, and a reference RT60-regression baseline. A Random Forest on lightweight time/spectral features reaches ~0.013 s MAE / ~0.022 s RMSE on a 36k/4k split. A streaming subset (1,000 linear + 3,000 circular array RIRs) is on Hugging Face; the full 50,000-RIR archive is on Zenodo.
✓ Dataset✓ Benchmark✓ Code
RIRDatasetRoom AcousticsRT60Reproducibility
RIR-Mega-Speech: A Reverberant Speech Corpus with Comprehensive Acoustic Metadata and Reproducible Evaluation
A large-scale reverberant speech corpus built by convolving LibriSpeech utterances with simulated RIRs from the RIR-Mega collection. Each reverberant utterance carries per-file acoustic metadata, enabling controlled analysis of reverberation effects on speech-processing systems with transparent, reproducible metrics.
✓ Dataset✓ Benchmark✓ Code
Reverberant SpeechASRCorpusReproducibility
Whisper-RIR-Mega: A Paired Clean-Reverberant Speech Benchmark for ASR Robustness to Room Acoustics
A benchmark of paired clean and reverberant speech for evaluating ASR robustness to room acoustics. Each sample pairs clean LibriSpeech audio with the same utterance convolved with a RIR-Mega impulse response, together with ground-truth transcripts and RIR metadata (RT60, DRR, C50). Accompanies a fine-tuned Whisper-Medium model specialized for reverberant speech.
✓ Dataset✓ Benchmark✓ Model✓ Code
BenchmarkWhisperASR RobustnessRoom Acoustics
Acoustivision Pro: An Open-Source Interactive Platform for Room Impulse Response Analysis and Acoustic Characterization
An open-source, interactive platform for visualizing and analyzing room impulse responses and characterizing acoustic spaces — bringing RT60, DRR, clarity (C50/C80), definition (D50), and early-decay-time analysis into an accessible, reproducible tool for the acoustics community.
✓ Code
ToolVisualizationAcoustic CharacterizationOpen Source
BeepBank-500: A Compact Synthetic Earcon / Alert Mini-Dataset for UI Sound Research
A fully synthetic earcon/alert mini-dataset of short tones and triads generated from a controlled parameter grid (waveform family, f0, duration, envelope, amplitude modulation, Schroeder-style reverbs). Ships with a metadata schema and lightweight baselines, released CC0 for UI-sound research.
✓ Dataset✓ Code
EarconsUI SoundSyntheticDataset