Browse State-of-the-Art › Speech-to-Text › Papers, page 4
Speech-to-Text
Papers archive 2025-07-28
archive papers tagged: 403 · with a code link: 129 · where Syntology ran a sample: 24 (18 with a run with no instrument failure, 6 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (24 of 403 tagged: 18 with a run with no instrument failure, 6 where every run was a failure of Syntology's instrument)
Page 4 of 5: papers 301 to 400 of 403, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Decision Attentive Regularization to Improve Simultaneous Speech Translation Systems13 Oct 2021 0 repositories listed
-
A Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation11 Oct 2021 0 repositories listed
-
Automated Testing of AI Models7 Oct 2021 0 repositories listed
-
Challenges and Opportunities of Speech Recognition for Bengali Language27 Sep 2021 0 repositories listed
-
Audio Interval Retrieval using Convolutional Neural Networks21 Sep 2021 0 repositories listed
-
Wav-BERT: Cooperative Acoustic and Linguistic Representation Learning for Low-Resource Speech Recognition19 Sep 2021 0 repositories listed
-
With One Voice: Composing a Travel Voice Assistant from Re-purposed Models4 Aug 2021 0 repositories listed
-
BTS: Back TranScription for Speech-to-Text Post-Processor using Text-to-Speech-to-Text1 Aug 2021 0 repositories listed
-
Corpus Creation and Evaluation for Speech-to-Text and Speech Translation1 Aug 2021 0 repositories listed
-
Multilingual Speech Translation from Efficient Finetuning of Pretrained Models1 Aug 2021 0 repositories listed
-
Improving Speech Translation by Understanding and Learning from the Auxiliary Text Translation Task12 Jul 2021 0 repositories listed
-
The USTC-NELSLIP Systems for Simultaneous Speech Translation Task at IWSLT 20211 Jul 2021 0 repositories listed
-
Pay Better Attention to Attention: Head Selection in Multilingual and Multi-Domain Sequence Modeling21 Jun 2021 0 repositories listed
-
Direct Simultaneous Speech-to-Text Translation Assisted by Synchronized Streaming ASR11 Jun 2021 0 repositories listed
-
10 Jun 2021 0 repositories listed
-
On the Design of Strategic Task Recommendations for Sustainable Crowdsourcing-Based Content Moderation4 Jun 2021 0 repositories listed
-
Findings of the Second Workshop on Automatic Simultaneous Translation1 Jun 2021 0 repositories listed
-
Worldly Wise (WoW) - Cross-Lingual Knowledge Fusion for Fact-based Visual Spoken-Question Answering1 Jun 2021 0 repositories listed
-
What shall we do with an hour of data? Speech recognition for the un- and under-served languages of Common Voice10 May 2021 0 repositories listed
-
A Benchmarking on Cloud based Speech-To-Text Services for French Speech and Background Noise Effect7 May 2021 0 repositories listed
-
Bridging the gap between streaming and non-streaming ASR systems bydistilling ensembles of CTC and RNN-T models25 Apr 2021 0 repositories listed
-
Label-Synchronous Speech-to-Text Alignment for ASR Using Forward and Backward Transformers21 Apr 2021 0 repositories listed
-
Multi-Discriminator Sobolev Defense-GAN Against Adversarial Attacks for End-to-End Speech Systems15 Mar 2021 0 repositories listed
-
Towards Robust Speech-to-Text Adversarial Attack15 Mar 2021 0 repositories listed
-
Towards the evaluation of automatic simultaneous speech translation from a communicative perspective15 Mar 2021 0 repositories listed
-
Inductive biases, pretraining and fine-tuning jointly account for brain responses to speech25 Feb 2021 0 repositories listed
-
NUVA: A Naming Utterance Verifier for Aphasia Treatment10 Feb 2021 0 repositories listed
-
Audio Adversarial Examples: Attacks Using Vocal Masks4 Feb 2021 0 repositories listed
-
Graph Neural Networks to Predict Customer Satisfaction Following Interactions with a Corporate Call Center31 Jan 2021 0 repositories listed
-
BCN2BRNO: ASR System Fusion for Albayzin 2020 Speech to Text Challenge29 Jan 2021 0 repositories listed
-
WER-BERT: Automatic WER Estimation with BERT in a Balanced Ordinal Classification Paradigm14 Jan 2021 0 repositories listed
-
15 Dec 2020 0 repositories listed
-
Incorporating Domain Knowledge To Improve Topic Segmentation Of Long MOOC Lecture Videos8 Dec 2020 0 repositories listed
-
A low latency ASR-free end to end spoken language understanding system10 Nov 2020 0 repositories listed
-
Effectively pretraining a speech translation decoder with Machine Translation data1 Nov 2020 0 repositories listed
-
Bridging the Modality Gap for Speech-to-Text Translation28 Oct 2020 0 repositories listed
-
Multilingual Speech Translation with Efficient Finetuning of Pretrained Models24 Oct 2020 0 repositories listed
-
Class-Conditional Defense GAN Against End-to-End Speech Attacks22 Oct 2020 0 repositories listed
-
MAM: Masked Acoustic Modeling for End-to-End Speech-to-Text Translation22 Oct 2020 0 repositories listed
-
A General Multi-Task Learning Framework to Leverage Text Data for Speech to Text Tasks21 Oct 2020 0 repositories listed
-
Ensemble Chinese End-to-End Spoken Language Understanding for Abnormal Event Detection from audio stream19 Oct 2020 0 repositories listed
-
Subtitles to Segmentation: Improving Low-Resource Speech-to-Text Translation Pipelines19 Oct 2020 0 repositories listed
-
Adversarial Attacks against Neural Networks in Audio Domain: Exploiting Principal Components14 Jul 2020 0 repositories listed
-
Contextualized Spoken Word Representations from Convolutional Autoencoders6 Jul 2020 0 repositories listed
-
1 Jul 2020 0 repositories listed
-
End-to-End Simultaneous Translation System for IWSLT2020 Using Modality Agnostic Meta-Learning1 Jul 2020 0 repositories listed
-
SimulSpeech: End-to-End Simultaneous Speech to Text Translation1 Jul 2020 0 repositories listed
-
Self-Supervised Representations Improve End-to-End Speech Translation22 Jun 2020 0 repositories listed
-
Exploration of End-to-End ASR for OpenSTT -- Russian Open Speech-to-Text Dataset15 Jun 2020 0 repositories listed
-
Improving Cross-Lingual Transfer Learning for End-to-End Speech Recognition with Speech Translation9 Jun 2020 0 repositories listed
-
ON-TRAC Consortium for End-to-End and Simultaneous Speech Translation Challenge Tasks at IWSLT 202024 May 2020 0 repositories listed
-
Speech to Text Adaptation: Towards an Efficient Cross-Modal Distillation17 May 2020 0 repositories listed
-
Crossing the SSH Bridge with Interview Data1 May 2020 0 repositories listed
-
SpiCE: A New Open-Access Corpus of Conversational Bilingual Speech in Cantonese and English1 May 2020 0 repositories listed
-
Subtitles to Segmentation: Improving Low-Resource Speech-to-TextTranslation Pipelines1 May 2020 0 repositories listed
-
Jointly Trained Transformers models for Spoken Language Translation25 Apr 2020 0 repositories listed
-
Cloud-Based Face and Speech Recognition for Access Control Applications23 Apr 2020 0 repositories listed
-
Learnings from Technological Interventions in a Low Resource Language: A Case-Study on Gondi21 Apr 2020 0 repositories listed
-
Speak2Label: Using Domain Knowledge for Creating a Large Scale Driver Gaze Zone Estimation Dataset13 Apr 2020 0 repositories listed
-
8 Apr 2020 0 repositories listed
-
A.I. based Embedded Speech to Text Using Deepspeech25 Feb 2020 0 repositories listed
-
A Comparative Study on End-to-end Speech to Text Translation20 Nov 2019 0 repositories listed
-
Data Efficient Direct Speech-to-Text Translation with Modality Agnostic Meta-Learning11 Nov 2019 0 repositories listed
-
8 Nov 2019 0 repositories listed
-
The IWSLT 2019 Evaluation Campaign1 Nov 2019 0 repositories listed
-
Analyzing ASR pretraining for low-resource speech-to-text translation23 Oct 2019 0 repositories listed
-
Instance-Based Model Adaptation For Direct Speech Translation23 Oct 2019 0 repositories listed
-
AeGAN: Time-Frequency Speech Denoising via Generative Adversarial Networks21 Oct 2019 0 repositories listed
-
DARTS: Dialectal Arabic Transcription System26 Sep 2019 0 repositories listed
-
Cross-lingual topic prediction for speech using translations29 Aug 2019 0 repositories listed
-
13 Aug 2019 0 repositories listed
-
Enhancing Transformer for End-to-end Speech-to-Text Translation1 Aug 2019 0 repositories listed
-
Analyzing Utility of Visual Context in Multimodal Speech Recognition Under Noisy Conditions30 Jun 2019 0 repositories listed
-
Telephonetic: Making Neural Language Models Robust to ASR and Semantic Noise13 Jun 2019 0 repositories listed
-
Natural Language Interactions in Autonomous Vehicles: Intent Detection and Slot Filling from Passenger Utterances23 Apr 2019 0 repositories listed
-
DeepCruiser: Automated Guided Testing for Stateful Deep Learning Systems13 Dec 2018 0 repositories listed
-
Development of Natural Language Processing Tools for Cook Islands Māori1 Dec 2018 0 repositories listed
-
A Voice Controlled E-Commerce Web Application16 Nov 2018 0 repositories listed
-
Leveraging Weakly Supervised Data to Improve End-to-End Speech-to-Text Translation5 Nov 2018 0 repositories listed
-
Towards Unsupervised Speech-to-Text Translation4 Nov 2018 0 repositories listed
-
Role of Intonation in Scoring Spoken English23 Aug 2018 0 repositories listed
-
Deep Learning Based Natural Language Processing for End to End Speech Translation9 Aug 2018 0 repositories listed
-
Unsupervised Cross-Modal Alignment of Speech and Text Embedding Spaces18 May 2018 0 repositories listed
-
Low-Resource Speech-to-Text Translation24 Mar 2018 0 repositories listed
-
Speech to text and text to speech recognition systems-Areview17 Mar 2018 0 repositories listed
-
The 2017 KIT IWSLT Speech-to-Text Systems for English and German1 Dec 2017 0 repositories listed
-
Visual Features for Context-Aware Speech Recognition1 Dec 2017 0 repositories listed
-
Interpreting Strategies Annotation in the WAW Corpus1 Sep 2017 0 repositories listed
-
Attention-Based End-to-End Speech Recognition on Voice Search22 Jul 2017 0 repositories listed
-
Polish Read Speech Corpus for Speech Tools and Services1 Jun 2017 0 repositories listed
-
Using of heterogeneous corpora for training of an ASR system1 Jun 2017 0 repositories listed
-
Towards speech-to-text translation without speech recognition13 Feb 2017 0 repositories listed
-
The 2016 KIT IWSLT Speech-to-Text Systems for English and German1 Dec 2016 0 repositories listed
-
Numerically Grounded Language Models for Semantic Error Correction14 Aug 2016 0 repositories listed
-
A Dutch Dysarthric Speech Database for Individualized Speech Therapy Research1 May 2016 0 repositories listed
-
LIA-RAG: a system based on graphs and divergence of probabilities applied to Speech-To-Text Summarization26 Jan 2016 0 repositories listed
-
The USFD Spoken Language Translation System for IWSLT 201413 Sep 2015 0 repositories listed
-
Voice based self help System: User Experience Vs Accuracy7 Apr 2015 0 repositories listed
-
Speech Recognition Web Services for Dutch1 May 2014 0 repositories listed
-
Noise in Speech-to-Text Voice: Analysis of Errors and Feasibility of Phonetic Similarity for Their Correction1 Dec 2013 0 repositories listed