Who Cited It

Adaptive Subgradient Methods for Online Learning and Stochastic Optimization

2010 · 8,621 citations · 37 from inside this corpus

John C. Duchi, Elad Hazan, Yoram Singer

The source holds an abstract for this work, but its best open-access copy is under no open licence, which does not permit us to republish the text. Read it at the source below.

Adaptive Subgradient Methods for Online Learning and Stochastic Optimization (2010)Adaptive Subgradient Methods …Term-weighting approaches in automatic text retrieval (1988)Term-weighting approaches in …RCV1: A New Benchmark Collection for Text Categorization Research (2004)RCV1: A New Benchmark Collect…Online Passive-Aggressive Algorithms (2006)Online Passive-Aggressive Alg…Adam: A Method for Stochastic Optimization (2014)Adam: A Method for Stochastic…Glove: Global Vectors for Word Representation (2014)Glove: Global Vectors for Wor…Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Sh… (2015)Batch Normalization: Accelera…Deep learning in neural networks: An overview (2014)Deep learning in neural netwo…Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Sh… (2024)Batch Normalization: Accelera…Convolutional Neural Networks for Sentence Classification (2014)Convolutional Neural Networks…Efficient Estimation of Word Representations in Vector Space (2013)Efficient Estimation of Word …Recursive Deep Models for Semantic Compositionality Over a Sentiment Treebank (2013)Recursive Deep Models for Sem…Deep Learning with Differential Privacy (2016)Deep Learning with Differenti…Advances and Open Problems in Federated Learning (2020)Advances and Open Problems in…An overview of gradient descent optimization algorithms (2016)An overview of gradient desce…Get To The Point: Summarization with Pointer-Generator Networks (2017)Get To The Point: Summarizati…On the difficulty of training Recurrent Neural Networks (2012)On the difficulty of training…A Convolutional Neural Network for Modelling Sentences (2014)A Convolutional Neural Networ…Deep Convolutional Neural Networks for Image Classification: A Comprehensive Review (2017)Deep Convolutional Neural Net…Knowledge Graph Embedding: A Survey of Approaches and Applications (2017)Knowledge Graph Embedding: A …Convolutional 2D Knowledge Graph Embeddings (2018)Convolutional 2D Knowledge Gr…Optimization as a Model for Few-Shot Learning (2017)Optimization as a Model for F…Attention-based LSTM for Aspect-level Sentiment Classification (2016)Attention-based LSTM for Aspe…MultiResUNet : Rethinking the U-Net architecture for multimodal biomedical image segmenta… (2019)MultiResUNet : Rethinking the…Privacy-Preserving Deep Learning (2015)Privacy-Preserving Deep Learn…Automatic differentiation in machine learning: a survey (2017)Automatic differentiation in …DeViSE: A Deep Visual-Semantic Embedding Model (2013)DeViSE: A Deep Visual-Semanti…Embedding Entities and Relations for Learning and Inference in Knowledge Bases (2014)Embedding Entities and Relati…Practical Recommendations for Gradient-Based Training of Deep Architectures (2012)Practical Recommendations for…Federated Learning with Non-IID Data (2018)Federated Learning with Non-I…A Fast and Accurate Dependency Parser using Neural Networks (2014)A Fast and Accurate Dependenc…Review of Deep Learning Algorithms and Architectures (2019)Review of Deep Learning Algor…Asynchronous Methods for Deep Reinforcement Learning (2016)Privacy-Preserving Deep Learning via Additively Homomorphic Encryption (2017)Privacy-Preserving Deep Learn…A Gift from Knowledge Distillation: Fast Optimization, Network Minimization and Transfer … (2017)A Gift from Knowledge Distill…Semantic Parsing on Freebase from Question-Answer Pairs (2013)Semantic Parsing on Freebase …Making Deep Neural Networks Robust to Label Noise: A Loss Correction Approach (2017)Making Deep Neural Networks R…
36 of 39 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

What cites it, inside the corpus

PaperYearCited
Adam: A Method for Stochastic Optimization201484,698
Glove: Global Vectors for Word Representation201434,067
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Sh…201524,404
Deep learning in neural networks: An overview201418,236
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Sh…202415,663
Convolutional Neural Networks for Sentence Classification201413,987
Efficient Estimation of Word Representations in Vector Space201311,714
Recursive Deep Models for Semantic Compositionality Over a Sentiment Treebank20136,849
Deep Learning with Differential Privacy20166,220
Advances and Open Problems in Federated Learning20205,383
An overview of gradient descent optimization algorithms20164,807
Get To The Point: Summarization with Pointer-Generator Networks20173,935
On the difficulty of training Recurrent Neural Networks20123,801
A Convolutional Neural Network for Modelling Sentences20143,571
Deep Convolutional Neural Networks for Image Classification: A Comprehensive Review20173,570
Knowledge Graph Embedding: A Survey of Approaches and Applications20172,703
Convolutional 2D Knowledge Graph Embeddings20182,467
Optimization as a Model for Few-Shot Learning20172,436
Attention-based LSTM for Aspect-level Sentiment Classification20162,380
MultiResUNet : Rethinking the U-Net architecture for multimodal biomedical image segmenta…20192,334
Privacy-Preserving Deep Learning20152,325
Automatic differentiation in machine learning: a survey20172,094
DeViSE: A Deep Visual-Semantic Embedding Model20132,068
Embedding Entities and Relations for Learning and Inference in Knowledge Bases20142,051
Practical Recommendations for Gradient-Based Training of Deep Architectures20121,960
Federated Learning with Non-IID Data20181,918
A Fast and Accurate Dependency Parser using Neural Networks20141,884
Review of Deep Learning Algorithms and Architectures20191,878
Asynchronous Methods for Deep Reinforcement Learning20161,689
Privacy-Preserving Deep Learning via Additively Homomorphic Encryption20171,665
A Gift from Knowledge Distillation: Fast Optimization, Network Minimization and Transfer …20171,660
Semantic Parsing on Freebase from Question-Answer Pairs20131,618
Making Deep Neural Networks Robust to Label Noise: A Loss Correction Approach20171,464

Topics

Stochastic Gradient Optimization TechniquesComputer Science
Advanced Bandit Algorithms ResearchDecision Sciences
Sparse and Compressive Sensing TechniquesEngineering

Is this record sound?

complete

Nothing in this record contradicts itself and no field we check is missing.

  • supports3 author record(s) attached.
  • supports47 reference(s) recorded.
  • neutralThe DOI carries no year to check against.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:43+00:00.

sha256 5cad55ac5d4d41e1…