Research Library
Library
Papers, patents & whitepapers, annotated
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren et al. · CVPR · 2015
Residual connections are in every transformer block you will ever use, which makes this required background even though the paper is about images. The framing is the lesson: the obstacle to depth was trainability, not capacity.
Computing numeric representations of words in a high-dimensional space
Google LLC · USPTO — US9037464B1 · 2015
The word-embedding idea as granted intellectual property. Useful for seeing how a technique that became foundational infrastructure was framed in claim language — and a reminder that the paper and the patent are separate artefacts with different purposes.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma, Jimmy Ba · ICLR · 2014
Still the default optimiser for essentially every model on this shelf, more than a decade on. Worth reading precisely because it is the piece of the stack most people never look inside.
Sequence to Sequence Learning with Neural Networks
Ilya Sutskever, Oriol Vinyals, Quoc V. Le · NIPS · 2014
The structural ancestor of every generative model you use. Its weakness — squeezing a whole sentence through one vector — is the specific problem attention was invented to solve three years later.
Generative Adversarial Networks
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza et al. · arXiv · 2014
The adversarial framing dominated generative modelling for most of a decade before diffusion displaced it. Worth reading for the idea that a learned critic can replace a hand-specified loss — a pattern that reappears throughout modern alignment work.
Efficient Estimation of Word Representations in Vector Space
Tomas Mikolov, Kai Chen, Greg Corrado et al. · ICLR · 2013
Where the idea that meaning can live in a vector became practical, and the direct ancestor of every embedding in your retrieval stack. Also the most approachable paper here — the model is simple enough to read in one sitting.
Showing 61–66 of 66 sources · newest first
About Nybble™
The AI space moves fast.
Nybble™ is how you keep up — and stay sharp.
What happened. In two minutes.
The AI news cycle moves at a pace no one can keep up with. Snack distills what launched, what shipped, and what matters — every day, without the filler.
Go to SnackThe concepts behind the headlines.
News tells you what. Stack tells you why and how. From RAG architectures to agentic evals, these are the ideas that will shape what you build next.
Go to StackProve you actually get it.
Reading about LangChain is not the same as knowing it. Hack challenges you with production-grade questions, then shows you the references that make the answer stick.
Go to Hack