Plain-language guides to the models, labs, and ideas shaping AI — synthesized from the corpus.
LoRAConcept
ReActConcept
GRPOConcept
AdamWConcept
Preparedness FrameworkConcept
MCPConcept
Proximal Policy OptimizationConcept
GRPO (Group Relative Policy Optimization)Concept
Differential PrivacyConcept
Reinforcement LearningConcept
Mixture of ExpertsConcept
Constitutional AIConcept
Chain-of-Thought ReasoningConcept
Diffusion ModelsConcept
Direct Preference Optimization (DPO)Concept
knowledge distillationConcept
LLM-as-a-JudgeConcept
mechanistic interpretabilityConcept
MambaConcept
PPOConcept
prompt injectionConcept
Model Context ProtocolConcept
Reinforcement Learning from Human FeedbackConcept
Reinforcement Learning with Verifiable RewardsConcept
Retrieval-Augmented GenerationConcept
scalable oversightConcept
supervised fine-tuningConcept
speculative decodingConcept
Vision-Language ModelsConcept