What does it really mean for an audio model to be robust?
Short, noisy, reverberant, filtered audio; real hardware.
Audio ML researcher and engineer
Audio ML · DSP · machine listening · acoustic privacy · real-time audio · MIR
Updated 6 Sep 2026

Short, noisy, reverberant, filtered audio; real hardware.
Privacy-preserving datasets and release pipelines.
Harmonic mixing, library organisation, real-time analysis.
Robustness, privacy, machine listening, and music technology.
Recovery fine-tuning can make a heavily pruned text-to-audio model look largely restored when it is evaluated under conditions similar to those used for fine-tuning. But that apparent recovery weakens sharply for shorter clips or out-of-domain prompts. The result shows why compressed generative models should be evaluated across real operating conditions, not at a single benchmark setting.
AudioLDM · Diffusion · Structured pruning · CLAP · PyTorch
This research evaluates how audio-language models detect speech when the audio is short, noisy, reverberant, or filtered. It compares Qwen2-Audio-7B, Qwen2-Audio-7B with LoRA, Qwen3-Omni-30B, and Silero VAD. The best result came from Qwen2-Audio-7B with LoRA and OPRO-Template: 93.3% balanced accuracy on 21,340 degraded clips.
Qwen · LoRA · OPRO · Silero · PyTorch
Sounds of Home is a residential audio dataset for sound event detection. It contains 1,344 one-hour recordings collected from 8 participants in Belgium, using AudioMoth recorders placed in living rooms and kitchens. Speech was removed before release, and PANNs predictions were provided for the audio frames.
SED · AudioMoth · PANNs · Datasets
A framework for removing speech from audio recordings before they are shared or published. The system supports privacy-preserving release workflows while retaining non-speech acoustic information for sound event detection research.
PANNs · AST · Silero VAD · WebRTC VAD · Audio privacy
This music analysis system estimates how well two EDM tracks mix harmonically. It analyzes tracks, computes chroma features, converts them into Tonal Interval Vectors, compares harmonic compatibility, and suggests pitch shifts that can improve a mix. The work began as an MSc thesis and later became an ICWE 2022 publication.
Chroma · TIV · Essentia · librosa · EDM
An end-to-end MVP for comparing raw transcription with enhance-and-transcribe on pre-recorded audio. The backend persists jobs, audio artifacts, transcripts, and provider payloads, with FastAPI, Celery, PostgreSQL, Redis, MinIO, Docker Compose, metrics, tracing, Grafana, and CI. It is a reproducible engineering prototype, not a production-hardened service.
FastAPI · Celery · Docker · PostgreSQL · Redis
Traktor ML turns a local Techno and Tech House library into Traktor-ready playlists. The pipeline extracts MERT embeddings, separates stems with Demucs, reads BPM and key metadata with Essentia, clusters similar tracks, orders them for smoother transitions, and exports M3U playlists. The current V4 run processed 239 tracks and exported 14 playlists.
MERT · Demucs · Essentia · HDBSCAN · UMAP · Streamlit
Edge Audio Labs · Montevideo, Uruguay · Hybrid
University of Surrey · Remote · Research collaboration
University of Surrey · Guildford, UK · Research engineering
Webhelp · Barcelona, Spain · Enterprise support
KPMG · Barcelona, Spain · Technology audit
Ikatu · Montevideo, Uruguay · Embedded R&D
Ikatu · Montevideo, Uruguay · Engineering internship
Universitat Pompeu Fabra · Barcelona, Spain
60 ECTS · Computational models, audio engineering, perception, cognition, and interactive systems
Master thesis on harmonic compatibility for EDM mixing. Final thesis grade: 9/10.
Universidad de la República - Facultad de Ingeniería · Montevideo, Uruguay
Equivalent to 300 ECTS · Specialisation in electronics and signal processing
Bachelor thesis on autonomous mobile robots communicated by software-defined radio.
School Nº265 ‘Virgilio Scarabelli Alberti’ · Montevideo, Uruguay
Musical language, choral singing, guitar, and instrumental ensembles
Thomas Deacon; Jennifer Williams; Jason R. C. Nurse; Christopher Hicks; Gabriel Bibbó; Arshdeep Singh; Mark D. Plumbley
2025 AES International Conference on Artificial Intelligence and Machine Learning for Audio
Gabriel Bibbó; Arshdeep Singh; Thomas Deacon; Mark D. Plumbley
2025 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA), Tahoe City, CA, October 2025
Gabriel Bibbó; Craig Cieciura; Mark D. Plumbley
Proceedings of the 54th International Congress and Exposition on Noise Control Engineering, São Paulo, Brazil, August 2025
Rhys Burchett-Vass; Arshdeep Singh; Gabriel Bibbó; Mark D. Plumbley
2025 AES International Conference on Artificial Intelligence and Machine Learning for Audio
Thomas Deacon; Gabriel Bibbó; Arshdeep Singh; Mark D. Plumbley
Forum Acusticum / Euronoise 2025, Málaga, Spain, June 2025
Arshdeep Singh; Haohe Liu; Gabriel Bibbó; Thomas Deacon; Mark D. Plumbley
IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2025), Hyderabad, India, April 2025
Gabriel Bibbó; Arshdeep Singh; Mark D. Plumbley
INTER-NOISE and NOISE-CON Congress and Conference Proceedings, Nantes, France, August 2024
Gabriel Bibbó; Thomas Deacon; Arshdeep Singh; Mark D. Plumbley
8th International Workshop on Speech Processing in Everyday Environments (CHiME 2024), Kos Island, Greece, September 2024
Thomas Deacon; Gabriel Bibbó; Arshdeep Singh; Mark D. Plumbley
International Conference on Sound and Music Computing (SMC 2024), Porto, Portugal, July 2024
Gabriel Bibbó; Arshdeep Singh; Mark D. Plumbley
IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA 2023), New York, U.S.A, October 2023
Gabriel Bibbó; Ángel Faraldo
International Conference on Web Engineering (ICWE 2022), Bari, Italy, July 2022
Gabriel Bibbó
Master thesis, Universitat Pompeu Fabra, Barcelona, Spain, 2021. Supervisor: Ángel Faraldo.
Gabriel Bibbó; Mariana Gelós; Martín Randall
Bachelor thesis, Universidad de la República, Montevideo, Uruguay, 2017. Supervisors: Pablo Belzarena and Federico Larroca.
Software engineering prototype for algorithmic trading infrastructure, with market data ingestion, event processing, risk controls, historical simulation, persistence, API access, and monitoring. It has not been tested in production.
Python · Backtesting · Risk controls · Monitoring
Raspberry Pi demo for real-time sound event recognition. The system runs pre-trained neural networks on a low-cost edge device, exposes a web interface, and can send email notifications when selected AudioSet events are detected.
Raspberry Pi · AudioSet · Real-time inference · Email notifications
Mechanical tool designed during the pandemic to avoid touching shared surfaces directly.
Mechanical design · Prototyping · Product development
IoT handwashing device for industrial environments. The device used stainless steel, WiFi, cloud connectivity, IR/RFID sensors, and a 3-litre tank.
WiFi · IR/RFID · Cloud connectivity · Industrial hygiene
Mobile app project for booking appointments and reducing crowding during the pandemic. Gabriel worked as product owner and project lead.
Product ownership · Civic tech · Project management
Certificate GR656321771GB · Expires 03/2027
Project management techniques, risk mitigation, resource allocation, project planning, and monitoring.
DeepLearning.AI · Coursera · Completed
CNNs, RNNs, transformers, TensorFlow, inductive transfer, and optimisation.
Stanford University · Coursera · Completed
Supervised and unsupervised learning and best practices in machine learning.
AURA · Montevideo, Uruguay · Graduated
Ableton and analogue instruments.
Montevideo, Uruguay
Python · C/C++ · PyTorch · Hugging Face · PEFT · TorchAudio · librosa · Essentia · mido · scikit-learn · pandas · NumPy · SciPy · Flask · FastAPI · Streamlit · Docker · Git · Linux CLI · Bash · Slurm · Redis · Prometheus · Grafana · PostgreSQL · SQLite · MATLAB · Unreal Engine 5.4 · FMOD · VS Code
CNNs · Transformers · Audio-Language Models · LoRA Fine-tuning · 4-bit Quantization · Supervised and Self-supervised Learning · Evaluation Pipelines · Statistical Testing · Edge Deployment
Sound Event Detection · Voice Activity Detection · Pitch and Onset Detection · Music Information Retrieval · Digital Signal Processing · Real-Time Audio · Perceptual Evaluation · DAWs · Ableton · DJing · Electronic Music Production
Reproducible ML pipelines · Automated Audio Testing · Dataset Curation · Open-Source Development · MLOps practices · AI-assisted Development · Technical Writing · Interdisciplinary Collaboration
Spanish Native · English C1 (CEFR) · Portuguese A2 (CEFR)
IEEE Signal Processing Society (2025)
EPSRC AI for Sound · University of Surrey
Available for academic collaboration, consulting, contract research, and selected remote opportunities in audio ML, machine listening, DSP, and music technology.