| 801 |
Deep Speech Denoising with Vector Space Projections
Jeff Hetherly, Paul Gamble, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
0 |
8 years ago |
| 802 |
Speech Driven Backchannel Generation using Deep Q-Network for Enhancing Engagement in Human-Robot Interaction
Nusrah Hussain, Engin Erzin, ... (+2 more)
|
👻
Ghosted
|
cs.AI
|
0 |
7 years ago |
| 803 |
Investigation of Large-Margin Softmax in Neural Language Modeling
Jingjing Huo, Yingbo Gao, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
6 years ago |
| 804 |
Sentence level estimation of psycholinguistic norms using joint multidimensional annotations
Anil Ramakrishna, Shrikanth Narayanan
|
👻
Ghosted
|
cs.CL
|
0 |
6 years ago |
| 805 |
Deep F-measure Maximization for End-to-End Speech Understanding
Leda Sarı, Mark Hasegawa-Johnson
|
👻
Ghosted
|
eess.AS
|
0 |
6 years ago |
| 806 |
Coswara: A website application enabling COVID-19 screening by analysing respiratory sound samples and health symptoms
Debarpan Bhattacharya, Debottam Dutta, ... (+9 more)
|
👻
Ghosted
|
cs.HC
|
0 |
4 years ago |
| 807 |
What can Speech and Language Tell us About the Working Alliance in Psychotherapy
Sebastian P. Bayerl, Gabriel Roccabruna, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
0 |
4 years ago |
| 808 |
ASR Error Detection via Audio-Transcript entailment
Nimshi Venkat Meripo, Sandeep Konam
|
👻
Ghosted
|
cs.CL
|
0 |
4 years ago |
| 809 |
Chronological Self-Training for Real-Time Speaker Diarization
Dirk Padfield, Daniel J. Liebling
|
👻
Ghosted
|
cs.SD
|
0 |
4 years ago |
| 810 |
Application of Knowledge Distillation to Multi-task Speech Representation Learning
Mine Kerpicci, Van Nguyen, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
0 |
3 years ago |
| 811 |
An Investigation of Indian Native Language Phonemic Influences on L2 English Pronunciations
Shelly Jain, Priyanshi Pal, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
0 |
3 years ago |
| 812 |
Memory Augmented Lookup Dictionary based Language Modeling for Automatic Speech Recognition
Yukun Feng, Ming Tu, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
0 |
3 years ago |
| 813 |
HumanDiffusion: diffusion model using perceptual gradients
Yota Ueda, Shinnosuke Takamichi, ... (+3 more)
|
👻
Ghosted
|
cs.HC
|
0 |
3 years ago |
| 814 |
Spoken Word2Vec: Learning Skipgram Embeddings from Speech
Mohammad Amaan Sayeed, Hanan Aldarmaki
|
👻
Ghosted
|
cs.CL
|
0 |
2 years ago |
| 815 |
Improving Label Assignments Learning by Dynamic Sample Dropout Combined with Layer-wise Optimization in Speech Separation
Chenyang Gao, Yue Gu, Ivan Marsic
|
👻
Ghosted
|
cs.SD
|
0 |
2 years ago |
| 816 |
Towards Multilingual Audio-Visual Question Answering
Orchid Chetia Phukan, Priyabrata Mallick, ... (+4 more)
|
👻
Ghosted
|
cs.LG
|
0 |
2 years ago |
| 817 |
Detecting the terminality of speech-turn boundary for spoken interactions in French TV and Radio content
Rémi Uro, Marie Tahon, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
2 years ago |
| 818 |
Gender Representation in TV and Radio: Automatic Information Extraction methods versus Manual Analyses
David Doukhan, Lena Dodson, ... (+7 more)
|
👻
Ghosted
|
eess.AS
|
0 |
2 years ago |
| 819 |
Exploring compressibility of transformer based text-to-music (TTM) models
Vasileios Moschopoulos, Thanasis Kotsiopoulos, ... (+8 more)
|
👻
Ghosted
|
eess.AS
|
0 |
2 years ago |
| 820 |
Analyzing Speech Motor Movement using Surface Electromyography in Minimally Verbal Adults with Autism Spectrum Disorder
Wazeer Zulfikar, Nishat Protyasha, ... (+14 more)
|
👻
Ghosted
|
q-bio.NC
|
0 |
2 years ago |
| 821 |
A Toolkit for Joint Speaker Diarization and Identification with Application to Speaker-Attributed ASR
Giovanni Morrone, Enrico Zovato, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 822 |
Relation-based Counterfactual Data Augmentation and Contrastive Learning for Robustifying Natural Language Inference Models
Heerin Yang, Sseung-won Hwang, Jungmin So
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 823 |
Analyzing Multimodal Features of Spontaneous Voice Assistant Commands for Mild Cognitive Impairment Detection
Nana Lin, Youxiang Zhu, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 824 |
GhostRNN: Reducing State Redundancy in RNN with Cheap Operations
Hang Zhou, Xiaoxu Zheng, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 825 |
A Unit-based System and Dataset for Expressive Direct Speech-to-Speech Translation
Anna Min, Chenxu Hu, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 826 |
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
Owais Mujtaba Khanday, Pablo Rodroguez San Esteban, ... (+3 more)
|
👻
Ghosted
|
cs.HC
|
0 |
1 year ago |
| 827 |
MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing
Junjie Zheng, Zihao Chen, ... (+6 more)
|
👻
Ghosted
|
cs.MM
|
0 |
1 year ago |
| 828 |
REWIND: Speech Time Reversal for Enhancing Speaker Representations in Diffusion-based Voice Conversion
Ishan D. Biyani, Nirmesh J. Shah, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 829 |
LASPA: Language Agnostic Speaker Disentanglement with Prefix-Tuned Cross-Attention
Aditya Srinivas Menon, Raj Prakash Gohil, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
0 |
1 year ago |
| 830 |
Leveraging Information Retrieval to Enhance Spoken Language Understanding Prompts in Few-Shot Learning
Pierre Lepagnol, Sahar Ghannay, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 831 |
Conformer-based Ultrasound-to-Speech Conversion
Ibrahim Ibrahimov, Zainkó Csaba, Gábor Gosztolya
|
👻
Ghosted
|
cs.SD
|
0 |
1 year ago |
| 832 |
Phonetically-Augmented Discriminative Rescoring for Voice Search Error Correction
Christophe Van Gysel, Maggie Wu, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 833 |
Modeling Probabilistic Reduction using Information Theory and Naive Discriminative Learning
Anna Stein, Kevin Tang
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 834 |
Multimodal Fusion with Semi-Supervised Learning Minimizes Annotation Quantity for Modeling Videoconference Conversation Experience
Andrew Chang, Chenkai Hu, ... (+6 more)
|
👻
Ghosted
|
eess.AS
|
0 |
1 year ago |
| 835 |
Leveraging Unlabeled Audio-Visual Data in Speech Emotion Recognition using Knowledge Distillation
Varsha Pendyala, Pedro Morgado, William Sethares
|
👻
Ghosted
|
cs.LG
|
0 |
1 year ago |
| 836 |
ClaritySpeech: Dementia Obfuscation in Speech
Dominika Woszczyk, Ranya Aloufi, Soteris Demetriou
|
👻
Ghosted
|
cs.CL
|
0 |
1 year ago |
| 837 |
Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries
Minyoung Kim, Sehwan Park, ... (+2 more)
|
👻
Ghosted
|
cs.SD
|
0 |
1 year ago |
| 838 |
CabinSep: IR-Augmented Mask-Based MVDR for Real-Time In-Car Speech Separation with Distributed Heterogeneous Arrays
Runduo Han, Yanxin Hu, ... (+5 more)
|
👻
Ghosted
|
cs.SD
|
0 |
1 year ago |