Who Cited It

Deep Speech: Scaling up end-to-end speech recognition

2014 · arXiv (Cornell University) · 1,519 citations · 5 from inside this corpus

Bryan Catanzaro, Greg Diamos, Erich Elsen, Ryan Prenger, Sanjeev Satheesh, Shubho Sengupta, Adam Coates, Andrew Y. Ng

The source holds an abstract for this work, but its best open-access copy is under no open licence, which does not permit us to republish the text. Read it at the source below.

Deep Speech: Scaling up end-to-end speech recognition (2014)Deep Speech: Scaling up end-t…Bidirectional recurrent neural networks (1997)Bidirectional recurrent neura…Improving neural networks by preventing co-adaptation of feature detectors (2012)Improving neural networks by …Connectionist temporal classification (2006)Connectionist temporal classi…Deep Sparse Rectifier Neural Networks (2011)Deep Sparse Rectifier Neural …Kaldi Speech Recognition Toolkit (2024)Kaldi Speech Recognition Tool…On the importance of initialization and momentum in deep learning (2013)On the importance of initiali…Sequence to Sequence Learning with Neural Networks (2014)Sequence to Sequence Learning…Context-Dependent Pre-Trained Deep Neural Networks for Large-Vocabulary Speech Recognition (2011)Context-Dependent Pre-Trained…Deep Neural Networks for Acoustic Modeling in Speech Recognition (2012)Deep Neural Networks for Acou…Towards End-To-End Speech Recognition with Recurrent Neural Networks (2014)Towards End-To-End Speech Rec…Acoustic Modeling Using Deep Belief Networks (2011)Acoustic Modeling Using Deep …SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition (2019)SpecAugment: A Simple Data Au…Listen, attend and spell: A neural network for large vocabulary conversational speech rec… (2016)Listen, attend and spell: A n…Privacy-Preserving Deep Learning (2015)Privacy-Preserving Deep Learn…Deep Speech 2: End-to-End Speech Recognition in English and Mandarin (2015)Deep Speech 2: End-to-End Spe…Unsupervised Data Augmentation for Consistency Training (2019)Unsupervised Data Augmentatio…
16 of 16 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

What cites it, inside the corpus

Topics

Speech Recognition and SynthesisComputer Science
Speech and Audio ProcessingComputer Science
Music and Audio ProcessingComputer Science

Is this record sound?

partial

One field of this record is missing or disagrees with another. What is shown below is what the source publishes.

  • supports8 author record(s) attached.
  • supports42 reference(s) recorded.
  • weakensThe DOI names 2021 but the record dates this to 2,014. One of the two is about a different paper.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:58+00:00.

sha256 88a60cdbb94c6a13…