Foundational AI Concepts
Plain-language reference pages for the foundational research concepts behind the AI tools in the MSP stack β transformers, attention, BERT, large language models, in-context learning, foundation model
Why this section exists
The five papers behind modern AI
Year
Paper
Concept it introduced
Entity page
How the concepts connect
Attention βββΆ Transformer βββ¬βββΆ BERT (bidirectional, fine-tuned)
β
ββββΆ GPT-3 / LLMs (autoregressive, in-context learning)
β
ββββΆ Foundation Models (the category + its risks)
ββββΆ RLHF / InstructGPT (aligning models to humans)Where the papers disagree
Start here
Last updated
Was this helpful?