We gratefully acknowledge support from
the Simons Foundation and member institutions.

Sound

Authors and titles for recent submissions

[ total of 37 entries: 1-10 | 11-20 | 21-30 | 31-37 ]
[ showing 10 entries per page: fewer | more | all ]

Mon, 3 Jun 2024

[1]  arXiv:2405.20887 [pdf, other]
Title: On the Condition Monitoring of Bolted Joints through Acoustic Emission and Deep Transfer Learning: Generalization, Ordinal Loss and Super-Convergence
Subjects: Sound (cs.SD); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[2]  arXiv:2405.20884 [pdf, other]
Title: Effects of Dataset Sampling Rate for Noise Cancellation through Deep Learning
Comments: 16 pages, 8 pictures, 3 tables
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Audio and Speech Processing (eess.AS)
[3]  arXiv:2405.20410 (cross-list from cs.CL) [pdf, other]
Title: SeamlessExpressiveLM: Speech Language Model for Expressive Speech-to-Speech Translation with Chain-of-Thought
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Sound (cs.SD); Audio and Speech Processing (eess.AS)
[4]  arXiv:2405.20402 (cross-list from eess.AS) [pdf, other]
Title: Cross-Talk Reduction
Comments: in International Joint Conference on Artificial Intelligence (IJCAI), 2024
Subjects: Audio and Speech Processing (eess.AS); Sound (cs.SD); Signal Processing (eess.SP)

Fri, 31 May 2024 (showing first 6 of 11 entries)

[5]  arXiv:2405.20289 [pdf, other]
Title: DITTO-2: Distilled Diffusion Inference-Time T-Optimization for Music Generation
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[6]  arXiv:2405.20172 [pdf, other]
Title: Iterative Feature Boosting for Explainable Speech Emotion Recognition
Comments: Published in: 2023 International Conference on Machine Learning and Applications (ICMLA)
Journal-ref: 2023 International Conference on Machine Learning and Applications (ICMLA), Jacksonville, FL, USA, 2023, pp. 543-549
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG); Audio and Speech Processing (eess.AS)
[7]  arXiv:2405.20101 [pdf, other]
Title: Fill in the Gap! Combining Self-supervised Representation Learning with Neural Audio Synthesis for Speech Inpainting
Subjects: Sound (cs.SD); Computation and Language (cs.CL); Audio and Speech Processing (eess.AS)
[8]  arXiv:2405.20059 [pdf, other]
Title: Spectral Mapping of Singing Voices: U-Net-Assisted Vocal Segmentation
Authors: Adam Sorrenti
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Audio and Speech Processing (eess.AS)
[9]  arXiv:2405.19796 [pdf, other]
Title: Explainable Attribute-Based Speaker Verification
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Audio and Speech Processing (eess.AS)
[10]  arXiv:2405.19343 [pdf, other]
Title: Luganda Speech Intent Recognition for IoT Applications
Comments: Presented as a conference paper at ICLR 2024/AfricaNLP
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Audio and Speech Processing (eess.AS)
[ total of 37 entries: 1-10 | 11-20 | 21-30 | 31-37 ]
[ showing 10 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, cs, new, 2406, contact, help  (Access key information)