Picture for Mark Cartwright

Mark Cartwright

New Jersey Institute of Technology

An Evaluation Framework for Structured Audio Captions Validated by Controlled Perturbations

Add code
Jul 23, 2026
Viaarxiv icon

Expressive Range Characterization of Open Text-to-Audio Models

Add code
Oct 31, 2025
Figure 1 for Expressive Range Characterization of Open Text-to-Audio Models
Figure 2 for Expressive Range Characterization of Open Text-to-Audio Models
Figure 3 for Expressive Range Characterization of Open Text-to-Audio Models
Figure 4 for Expressive Range Characterization of Open Text-to-Audio Models
Viaarxiv icon

EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation

Add code
Oct 15, 2024
Figure 1 for EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
Figure 2 for EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
Figure 3 for EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
Figure 4 for EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
Viaarxiv icon

Compositional Audio Representation Learning

Add code
Sep 15, 2024
Figure 1 for Compositional Audio Representation Learning
Figure 2 for Compositional Audio Representation Learning
Figure 3 for Compositional Audio Representation Learning
Figure 4 for Compositional Audio Representation Learning
Viaarxiv icon

Multi-label Open-set Audio Classification

Add code
Oct 20, 2023
Figure 1 for Multi-label Open-set Audio Classification
Figure 2 for Multi-label Open-set Audio Classification
Figure 3 for Multi-label Open-set Audio Classification
Figure 4 for Multi-label Open-set Audio Classification
Viaarxiv icon

A General Framework for Learning Procedural Audio Models of Environmental Sounds

Add code
Mar 04, 2023
Viaarxiv icon

Urban Rhapsody: Large-scale exploration of urban soundscapes

Add code
May 25, 2022
Figure 1 for Urban Rhapsody: Large-scale exploration of urban soundscapes
Figure 2 for Urban Rhapsody: Large-scale exploration of urban soundscapes
Figure 3 for Urban Rhapsody: Large-scale exploration of urban soundscapes
Figure 4 for Urban Rhapsody: Large-scale exploration of urban soundscapes
Viaarxiv icon

A Study on Robustness to Perturbations for Representations of Environmental Sound

Add code
Mar 23, 2022
Figure 1 for A Study on Robustness to Perturbations for Representations of Environmental Sound
Figure 2 for A Study on Robustness to Perturbations for Representations of Environmental Sound
Figure 3 for A Study on Robustness to Perturbations for Representations of Environmental Sound
Figure 4 for A Study on Robustness to Perturbations for Representations of Environmental Sound
Viaarxiv icon

Who calls the shots? Rethinking Few-Shot Learning for Audio

Add code
Oct 18, 2021
Figure 1 for Who calls the shots? Rethinking Few-Shot Learning for Audio
Figure 2 for Who calls the shots? Rethinking Few-Shot Learning for Audio
Figure 3 for Who calls the shots? Rethinking Few-Shot Learning for Audio
Figure 4 for Who calls the shots? Rethinking Few-Shot Learning for Audio
Viaarxiv icon

Weakly Supervised Source-Specific Sound Level Estimation in Noisy Soundscapes

Add code
May 06, 2021
Figure 1 for Weakly Supervised Source-Specific Sound Level Estimation in Noisy Soundscapes
Figure 2 for Weakly Supervised Source-Specific Sound Level Estimation in Noisy Soundscapes
Figure 3 for Weakly Supervised Source-Specific Sound Level Estimation in Noisy Soundscapes
Viaarxiv icon