| 701 |
Machine Assisted Analysis of Vowel Length Contrasts in Wolof
Elodie Gauthier, Laurent Besacier, Sylvie Voisin
|
👻
Ghosted
|
cs.CL
|
3 |
9 years ago |
| 702 |
Singing voice phoneme segmentation by hierarchically inferring syllable and phoneme onset positions
Rong Gong, Xavier Serra
|
👻
Ghosted
|
cs.SD
|
3 |
8 years ago |
| 703 |
Active Annotation: bootstrapping annotation lexicon and guidelines for supervised NLU learning
Federico Marinelli, Alessandra Cervone, ... (+4 more)
|
👻
Ghosted
|
cs.CL
|
3 |
7 years ago |
| 704 |
An End-to-End Audio Classification System based on Raw Waveforms and Mix-Training Strategy
Jiaxu Chen, Jing Hao, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
3 |
6 years ago |
| 705 |
Evaluating Automatically Generated Phoneme Captions for Images
Justin van der Hout, Zoltán D'Haese, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
3 |
6 years ago |
| 706 |
"This is Houston. Say again, please". The Behavox system for the Apollo-11 Fearless Steps Challenge (phase II)
Arseniy Gorin, Daniil Kulko, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
3 |
6 years ago |
| 707 |
Automatic Quality Assessment for Audio-Visual Verification Systems. The LOVe submission to NIST SRE Challenge 2019
Grigory Antipov, Nicolas Gengembre, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
3 |
6 years ago |
| 708 |
Complementary Language Model and Parallel Bi-LRNN for False Trigger Mitigation
Rishika Agarwal, Xiaochuan Niu, ... (+4 more)
|
👻
Ghosted
|
eess.AS
|
3 |
6 years ago |
| 709 |
AutoSpeech 2020: The Second Automated Machine Learning Challenge for Speech Classification
Jingsong Wang, Tom Ko, ... (+5 more)
|
👻
Ghosted
|
cs.AI
|
3 |
5 years ago |
| 710 |
Augmenting Images for ASR and TTS through Single-loop and Dual-loop Multimodal Chain Framework
Johanes Effendi, Andros Tjandra, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
3 |
5 years ago |
| 711 |
Text Augmentation for Language Models in High Error Recognition Scenario
Karel Beneš, Lukáš Burget
|
👻
Ghosted
|
cs.CL
|
3 |
5 years ago |
| 712 |
A Temporal Extension of Latent Dirichlet Allocation for Unsupervised Acoustic Unit Discovery
Werner van der Merwe, Herman Kamper, Johan du Preez
|
👻
Ghosted
|
eess.AS
|
3 |
4 years ago |
| 713 |
Non-Linear Pairwise Language Mappings for Low-Resource Multilingual Acoustic Model Fusion
Muhammad Umar Farooq, Darshan Adiga Haniya Narayana, Thomas Hain
|
👻
Ghosted
|
cs.CL
|
3 |
4 years ago |
| 714 |
Multiple-hypothesis RNN-T Loss for Unsupervised Fine-tuning and Self-training of Neural Transducer
Cong-Thanh Do, Mohan Li, Rama Doddipatla
|
👻
Ghosted
|
cs.CL
|
3 |
4 years ago |
| 715 |
VQ-T: RNN Transducers using Vector-Quantized Prediction Network States
Jiatong Shi, George Saon, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
3 |
4 years ago |
| 716 |
Enhancing Semantic Understanding with Self-supervised Methods for Abstractive Dialogue Summarization
Hyunjae Lee, Jaewoong Yun, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
3 |
4 years ago |
| 717 |
Integrating Form and Meaning: A Multi-Task Learning Model for Acoustic Word Embeddings
Badr M. Abdullah, Bernd Möbius, Dietrich Klakow
|
👻
Ghosted
|
cs.CL
|
3 |
3 years ago |
| 718 |
From Disfluency Detection to Intent Detection and Slot Filling
Mai Hoang Dao, Thinh Hung Truong, Dat Quoc Nguyen
|
👻
Ghosted
|
cs.CL
|
3 |
3 years ago |
| 719 |
Extending Compositional Attention Networks for Social Reasoning in Videos
Christina Sartzetaki, Georgios Paraskevopoulos, Alexandros Potamianos
|
👻
Ghosted
|
cs.CV
|
3 |
3 years ago |
| 720 |
Pronunciation Modeling of Foreign Words for Mandarin ASR by Considering the Effect of Language Transfer
Lei Wang, Rong Tong
|
👻
Ghosted
|
cs.CL
|
3 |
3 years ago |
| 721 |
Joint Speech Translation and Named Entity Recognition
Marco Gaido, Sara Papi, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
3 |
3 years ago |
| 722 |
Random Utterance Concatenation Based Data Augmentation for Improving Short-video Speech Recognition
Yist Y. Lin, Tao Han, ... (+7 more)
|
👻
Ghosted
|
eess.AS
|
3 |
3 years ago |
| 723 |
Blank Collapse: Compressing CTC emission for the faster decoding
Minkyu Jung, Ohhyeok Kwon, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
3 |
3 years ago |
| 724 |
Generating Multilingual Gender-Ambiguous Text-to-Speech Voices
Konstantinos Markopoulos, Georgia Maniati, ... (+10 more)
|
👻
Ghosted
|
cs.SD
|
3 |
3 years ago |
| 725 |
Quantifying the perceptual value of lexical and non-lexical channels in speech
Sarenne Wallbridge, Peter Bell, Catherine Lai
|
👻
Ghosted
|
cs.CL
|
3 |
3 years ago |
| 726 |
Video Multimodal Emotion Recognition System for Real World Applications
Sun-Kyung Lee, Jong-Hwan Kim
|
👻
Ghosted
|
cs.HC
|
3 |
3 years ago |
| 727 |
Modular Speech-to-Text Translation for Zero-Shot Cross-Modal Transfer
Paul-Ambroise Duquenne, Holger Schwenk, Benoît Sagot
|
👻
Ghosted
|
cs.CL
|
3 |
2 years ago |
| 728 |
Voice Passing : a Non-Binary Voice Gender Prediction System for evaluating Transgender voice transition
David Doukhan, Simon Devauchelle, ... (+5 more)
|
👻
Ghosted
|
eess.AS
|
3 |
2 years ago |
| 729 |
Resource-Efficient Speech Quality Prediction through Quantization Aware Training and Binary Activation Maps
Mattias Nilsson, Riccardo Miccini, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
3 |
2 years ago |
| 730 |
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
Ailin Liu, Pepijn Vunderink, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
3 |
2 years ago |
| 731 |
Zero-Shot Mono-to-Binaural Speech Synthesis
Alon Levkovitch, Julian Salazar, ... (+5 more)
|
👻
Ghosted
|
cs.SD
|
3 |
1 year ago |
| 732 |
ADI-20: Arabic Dialect Identification dataset and models
Haroun Elleuch, Salima Mdhaffar, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
3 |
9 months ago |
| 733 |
Automatic Pronunciation Generation by Utilizing a Semi-supervised Deep Neural Networks
Naoya Takahashi, Tofigh Naghibi, Beat Pfister
|
👻
Ghosted
|
cs.CL
|
2 |
10 years ago |
| 734 |
The Trajectory of Voice Onset Time with Vocal Aging
Xuanda Chen, Ziyu Xiong, Jian Hu
|
👻
Ghosted
|
cs.SD
|
2 |
7 years ago |
| 735 |
Exploring Methods for the Automatic Detection of Errors in Manual Transcription
Xiaofei Wang, Jinyi Yang, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 736 |
Performance Monitoring for End-to-End Speech Recognition
Ruizhi Li, Gregory Sell, Hynek Hermansky
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 737 |
M2H-GAN: A GAN-based Mapping from Machine to Human Transcripts for Speech Understanding
Titouan Parcollet, Mohamed Morchid, ... (+2 more)
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 738 |
Sampling from Stochastic Finite Automata with Applications to CTC Decoding
Martin Jansche, Alexander Gutkin
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 739 |
Large-Scale Speaker Diarization of Radio Broadcast Archives
Emre Yılmaz, Adem Derinel, ... (+5 more)
|
👻
Ghosted
|
cs.CL
|
2 |
7 years ago |
| 740 |
Avaya Conversational Intelligence: A Real-Time System for Spoken Language Understanding in Human-Human Call Center Conversations
Jan Mizgajski, Adrian Szymczak, ... (+15 more)
|
👻
Ghosted
|
eess.AS
|
2 |
7 years ago |
| 741 |
Reverse Transfer Learning: Can Word Embeddings Trained for Different NLP Tasks Improve Neural Language Models?
Lyan Verwimp, Jerome R. Bellegarda
|
👻
Ghosted
|
cs.CL
|
2 |
6 years ago |
| 742 |
Weakly Supervised Training of Hierarchical Attention Networks for Speaker Identification
Yanpei Shi, Qiang Huang, Thomas Hain
|
👻
Ghosted
|
eess.AS
|
2 |
6 years ago |
| 743 |
Efficient MDI Adaptation for n-gram Language Models
Ruizhe Huang, Ke Li, ... (+3 more)
|
👻
Ghosted
|
cs.CL
|
2 |
6 years ago |
| 744 |
Emotion Carrier Recognition from Personal Narratives
Aniruddha Tammewar, Alessandra Cervone, Giuseppe Riccardi
|
👻
Ghosted
|
cs.CL
|
2 |
6 years ago |
| 745 |
Pardon the Interruption: An Analysis of Gender and Turn-Taking in U.S. Supreme Court Oral Arguments
Haley Lepp, Gina-Anne Levow
|
👻
Ghosted
|
cs.CL
|
2 |
5 years ago |
| 746 |
Masked Proxy Loss For Text-Independent Speaker Verification
Jiachen Lian, Aiswarya Vinod Kumar, ... (+3 more)
|
👻
Ghosted
|
cs.SD
|
2 |
5 years ago |
| 747 |
Meta Auxiliary Learning for Low-resource Spoken Language Understanding
Yingying Gao, Junlan Feng, ... (+2 more)
|
👻
Ghosted
|
eess.AS
|
2 |
4 years ago |
| 748 |
Introducing Auxiliary Text Query-modifier to Content-based Audio Retrieval
Daiki Takeuchi, Yasunori Ohishi, ... (+3 more)
|
👻
Ghosted
|
eess.AS
|
2 |
4 years ago |
| 749 |
Deep Sparse Conformer for Speech Recognition
Xianchao Wu
|
👻
Ghosted
|
cs.CL
|
2 |
4 years ago |
| 750 |
SF-DST: Few-Shot Self-Feeding Reading Comprehension Dialogue State Tracking with Auxiliary Task
Jihyun Lee, Gary Geunbae Lee
|
👻
Ghosted
|
cs.CL
|
2 |
3 years ago |