Who Cited It

Pre-trained models for natural language processing: A survey

2020 · Science China Technological Sciences · 1,521 citations · 2 from inside this corpus

Xipeng Qiu, Tianxiang Sun, Yige Xu, Yunfan Shao low, Ning Dai, Xuanjing Huang

No abstract in the source record.

Pre-trained models for natural language processing: A survey (2020)Pre-trained models for natura…Long Short-Term Memory (1997)Long Short-Term MemoryAI-Assisted Pipeline for Dynamic Generation of Trustworthy Health Supplement Content at S… (2018)AI-Assisted Pipeline for Dyna…Glove: Global Vectors for Word Representation (2014)Glove: Global Vectors for Wor…BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding (2019)BERT: Pre-training of Deep Bi…A Survey on Transfer Learning (2009)A Survey on Transfer LearningReducing the Dimensionality of Data with Neural Networks (2006)Reducing the Dimensionality o…Distributed Representations of Words and Phrases and their Compositionality (2013)Distributed Representations o…HISTORIAE, History of Socio-Cultural Transformation as Linguistic Data Science. A Humanit… (2019)HISTORIAE, History of Socio-C…Detecting Functionality-Specific Vulnerabilities via Retrieving Individual Functionality-… (2025)Detecting Functionality-Speci…Distilling the Knowledge in a Neural Network (2015)Distilling the Knowledge in a…Convolutional Neural Networks for Sentence Classification (2014)Convolutional Neural Networks…Sequence to Sequence Learning with Neural Networks (2014)Sequence to Sequence Learning…Enriching Word Vectors with Subword Information (2017)Enriching Word Vectors with S…Explainable Artificial Intelligence (XAI): Concepts, taxonomies, opportunities and challe… (2019)Explainable Artificial Intell…Semi-Supervised Classification with Graph Convolutional Networks (2016)Semi-Supervised Classificatio…[No title in the source record — Edinburgh Research Explorer (University of Edinburgh)][No title in the source recor…Attention Is All You Need (2025)Attention Is All You NeedRecursive Deep Models for Semantic Compositionality Over a Sentiment Treebank (2013)Recursive Deep Models for Sem…SQuAD: 100,000+ Questions for Machine Comprehension of Text (2016)SQuAD: 100,000+ Questions for…Natural Language Processing (almost) from Scratch (2011)Natural Language Processing (…Distributed Representations of Sentences and Documents (2014)Distributed Representations o…DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter (2019)DistilBERT, a distilled versi…ALBERT: A Lite BERT for Self-supervised Learning of Language\n Representations (2019)ALBERT: A Lite BERT for Self-…GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding (2018)GLUE: A Multi-Task Benchmark …Neural Architecture Search with Reinforcement Learning (2016)Neural Architecture Search wi…Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer (2019)Exploring the Limits of Trans…A Convolutional Neural Network for Modelling Sentences (2014)A Convolutional Neural Networ…Sequence to Sequence Learning with Neural Networks (2014)Sequence to Sequence Learning…Transformer-XL: Attentive Language Models beyond a Fixed-Length Context (2019)Transformer-XL: Attentive Lan…SciBERT: A Pretrained Language Model for Scientific Text (2019)SciBERT: A Pretrained Languag…Linguistic Regularities in Continuous Space Word Representations (2013)Linguistic Regularities in Co…Proceedings of the International Joint Conference on Artificial Intelligence 2007 (2007)Proceedings of the Internatio…LXMERT: Learning Cross-Modality Encoder Representations from Transformers (2019)LXMERT: Learning Cross-Modali…Model compression (2006)Model compressionDeep Learning--based Text Classification (2021)Deep Learning--based Text Cla…A Brief Overview of ChatGPT: The History, Status Quo and Potential Future Development (2023)A Brief Overview of ChatGPT: …
36 of 38 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

PaperYearCited
Long Short-Term Memory1997101,359
AI-Assisted Pipeline for Dynamic Generation of Trustworthy Health Supplement Content at S…201846,036
Glove: Global Vectors for Word Representation201434,067
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding201933,416
A Survey on Transfer Learning200923,670
Reducing the Dimensionality of Data with Neural Networks200621,247
Distributed Representations of Words and Phrases and their Compositionality201318,054
HISTORIAE, History of Socio-Cultural Transformation as Linguistic Data Science. A Humanit…201917,489
Detecting Functionality-Specific Vulnerabilities via Retrieving Individual Functionality-…202516,314
Distilling the Knowledge in a Neural Network201514,099
Convolutional Neural Networks for Sentence Classification201413,987
Sequence to Sequence Learning with Neural Networks201413,351
Enriching Word Vectors with Subword Information20179,900
Explainable Artificial Intelligence (XAI): Concepts, taxonomies, opportunities and challe…20199,796
Semi-Supervised Classification with Graph Convolutional Networks20168,058
[No title in the source record — Edinburgh Research Explorer (University of Edinburgh)]7,291
Attention Is All You Need20257,133
Recursive Deep Models for Semantic Compositionality Over a Sentiment Treebank20136,849
SQuAD: 100,000+ Questions for Machine Comprehension of Text20166,435
Natural Language Processing (almost) from Scratch20115,174
Distributed Representations of Sentences and Documents20145,121
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter20194,600
ALBERT: A Lite BERT for Self-supervised Learning of Language\n Representations20194,076
GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding20184,055
Neural Architecture Search with Reinforcement Learning20163,881
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer20193,698
A Convolutional Neural Network for Modelling Sentences20143,571
Sequence to Sequence Learning with Neural Networks20143,514
Transformer-XL: Attentive Language Models beyond a Fixed-Length Context20193,202
SciBERT: A Pretrained Language Model for Scientific Text20193,109
Linguistic Regularities in Continuous Space Word Representations20132,885
Proceedings of the International Joint Conference on Artificial Intelligence 200720072,765
LXMERT: Learning Cross-Modality Encoder Representations from Transformers20192,335
Model compression20062,118

What cites it, inside the corpus

Topics

Topic ModelingComputer Science
Natural Language Processing TechniquesComputer Science
Multimodal Machine Learning ApplicationsComputer Science

Is this record sound?

complete

Nothing in this record contradicts itself and no field we check is missing.

  • supports6 author record(s) attached.
  • supports325 reference(s) recorded.
  • neutralThe DOI carries no year to check against.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:57+00:00.

sha256 b3024609427ad450…