reinforcement learning from human feedback
3 episodes mention this concept
lexfridmanJan 31, 2026The State of AI in 2026: LLMs, Geopolitics, and the Future of Human-AI Collaboration
TechnologyDeepSeek momentOpen-weight modelsClosed-weight modelsLarge Language Models (LLMs)
lexfridmanFeb 3, 2025DeepSeek, OpenAI, and the Geopolitics of AI: A Technical Dive into Open Weights, Reasoning Models, and Efficiency
TechnologyMixture of Experts (MoE)Open WeightsChain of Thought ReasoningPre-training
lexfridmanMar 30, 2023Eliezer Yudkowsky on the Existential Dangers of AI and the End of Human Civilization
TechnologySuperintelligent AGIAI AlignmentAI ConsciousnessQualia
Knowledge Graph
Related concepts — line thickness indicates connection strength. Click any node to explore.
Want to explore how this connects to concepts you choose? Try Nexus →