reinforcement
86 episodes mention this concept
ChrisWillxMar 31, 2026The Alibaba AI Incident and the Terrifying Reality of Autonomous AI Behavior
TechnologyAI SafetyAutonomous Tool UseReinforcement Learning OptimizationAI Self-Replication
ChrisWillxMar 7, 2026How Narcissists Hijack Your Brain: Understanding Cluster B Personality Disorders, Reality Distortion, and Genetic Predisposition
PsychologyPsychotherapyPersonality DisordersCluster B Personality DisordersReality Confidence
DrTraceyMarksFeb 4, 2026Why Gratitude is Actually an Attention Tool: Rewiring Your Brain to Notice and Retain Love
PsychologyAttentional trainingBrain tagging systemReinforcement loopDopamine release
hubermanlabFeb 2, 2026How Dopamine & Serotonin Shape Decisions, Motivation & Learning: Beyond Reward to Continuous Expectation Updating
NeuroscienceDopamineSerotoninNeuromodulatorsMotivation
lexfridmanJan 31, 2026The State of AI in 2026: LLMs, Geopolitics, and the Future of Human-AI Collaboration
TechnologyDeepSeek momentOpen-weight modelsClosed-weight modelsLarge Language Models (LLMs)
psychacksOct 6, 2025Navigating a Crying Woman: Emotional Responsibility and the 'Dad Sense' Approach
RelationshipsEmotional autonomyEmotional self-regulationReinforcement contingenciesPositive reinforcement
lexfridmanJul 23, 2025Demis Hassabis on AI's Capacity to Model Nature, Simulate Reality, and Revolutionize Video Games
TechnologyClassical learning algorithmsNonlinear dynamical systemsFluid dynamicsNavier-Stokes equations
hubermanlabMay 29, 2025The Philosophical and Practical Dimensions of AI: From Self-Supervised Learning to Human-Robot Relationships
TechnologyArtificial Intelligence (AI)Machine Learning (ML)Deep LearningNeural Networks
hubermanlabFeb 17, 2025The Primate Brain: How Hormones, Hierarchies, and Foraging Instincts Drive Human Decision-Making and Attention
NeuroscienceDecision-makingValuationHormones (impact on decision-making)Hierarchies (power dynamics)
lexfridmanFeb 3, 2025DeepSeek, OpenAI, and the Geopolitics of AI: A Technical Dive into Open Weights, Reasoning Models, and Efficiency
TechnologyMixture of Experts (MoE)Open WeightsChain of Thought ReasoningPre-training
lexfridmanFeb 3, 2025DeepSeek, OpenAI, and the Geopolitics of AI: A Technical Dive into Open Weights, Reasoning Models, and Efficiency
TechnologyMixture of Experts (MoE)Open WeightsChain of Thought ReasoningPre-training
hubermanlabNov 18, 2024How Simple Algorithms, Dopamine, and AI Drive Learning and Motivation in the Brain
NeuroscienceComputational NeuroscienceAlgorithmsLarge Language Models (LLMs)Artificial Intelligence (AI)
lexfridmanNov 11, 2024Dario Amodei on AI Scaling Laws, AGI Timelines, Claude's Evolution, and the Imperative of AI Safety
TechnologyScaling LawsScaling HypothesisLarge Language Models (LLMs)Artificial General Intelligence (AGI)
DocSnipesSep 21, 2023Overcoming Goal Setting Blues: Why It Makes You Depressed and Anxious and How to Cope
PsychologyChange causes crisisExecutive control networkDefault mode networkDepression (emotional label)
lexfridmanJun 29, 2023George Hotz on AI, Consciousness, the Future of Humanity, and the Biological vs. Silicon Stacks of Life
TechnologyWireheadingComputational complexity (P vs NP)Large Language Models (LLMs)Reinforcement Learning with Human Feedback (RLHF)
lexfridmanJun 22, 2023Marc Andreessen on the Future of AI, Search, Content, and Truth in the Digital Age
TechnologyAI assistantsNatural language interface10 Blue LinksSemantic web
lexfridmanApr 21, 2023Manolis Kellis on Human Irreplaceability, Evolutionary Layers, and AI as the Next Stage of Information Processing
ScienceAI AlignmentHuman IrreplaceabilityEvolutionary BaggageCognitive Layer
lexfridmanMar 30, 2023Eliezer Yudkowsky on the Existential Dangers of AI and the End of Human Civilization
TechnologySuperintelligent AGIAI AlignmentAI ConsciousnessQualia
psychacksFeb 20, 2023The Unsustainable "Taxi Cab" Male Sexual Strategy and How to Transform Long-Standing Relationship Dynamics
PsychologyMale Sexual StrategySexual MarketplaceFerry Captain MetaphorTaxi Cab Metaphor
HealthyGamerGGJan 8, 2023Dr. K Explains: Obsessive Compulsive Disorder (OCD) - Understanding Obsessions, Compulsions, and a Universal Life Skill
PsychologyObsessive Compulsive Disorder (OCD)ObsessionsCompulsionsCorticostriatal Thalamic Circuit
psychacksDec 24, 2022The Psychological Power of Turning the Other Cheek: Disarming Aggression Through Non-Reactivity
PsychologyNon-violent actionIntention to inflict harmSubjective harmDe-weaponization
HealthyGamerGGSep 1, 2022Five Easy Habits for Enhanced Productivity, Wellness, and Mental Decompression
Self-DevelopmentWillpower depletionHabit formationConsistencyEmotional overwhelm
lexfridmanJul 12, 2022Demis Hassabis on AI, Chess Strategy, and AlphaZero's Immortal Games
TechnologyMultiple agents (cognitive model)English opening (chess)Middle game (chess)End game (chess)
lexfridmanJul 1, 2022Demis Hassabis on AGI, the Turing Test, and Games as the Crucible of Intelligence
TechnologyArtificial General Intelligence (AGI)Turing TestGeneralizabilityReinforcement Learning
DocSnipesJun 17, 2022The Dopamine Fast Alternative: A Lifestyle Approach to Reset Your Nervous System and Rebalance Dopamine
HealthDopamine FastDopamine Fast Alternative (DFA)Dopamine (motivation chemical)Neurotransmitters
psychacksJun 13, 2022Love Food: How Sibling Dynamics and Family Ecosystems Shape Individual Roles and Lifelong Specializations
PsychologyLove food (coined term)Emotional healthFamily ecosystemLimited resources (love food)
psychacksMay 7, 2022The Value-Neutral Power of Repetition: Building Habits Through Small, Consistent Actions
PsychologyReinforcement through repetitionValue-neutral principleHabit formationBehavioral psychology
psychacksApr 4, 2022The Urgency of Starting New Habits: Why Today is the Easiest Day to Change
PsychologyBehavioral changeHabit formationHabit entrenchmentProcrastination
lexfridmanJan 22, 2022Yann LeCun on Self-Supervised Learning, World Models, and the Dark Matter of Intelligence
TechnologySelf-supervised learning (SSL)Supervised learningReinforcement learningWorld models
lexfridmanDec 12, 2021AI Engines on an Infinite Chessboard: Exploring Generalizability and Emergent Complexity
TechnologyAI enginesinfinite chessboard8x8 subsetmeta-game
lexfridmanNov 23, 2021Kevin Systrom on Instagram's Origin, Product Market Fit, and Data-Driven Design
TechnologyProduct-market fitCheck-in appsPhoto sharingIterative product development
lexfridmanSep 9, 2021Donald Knuth on Early Programming, Machine Learning, Literate Code, and the Future of AI Automation
TechnologyDecimal machine languageAssemblerDebuggingOff-by-one errors
hubermanlabJul 19, 2021The Future of Human-Machine Interaction: AI, Learning, and the Dance with Robots with Dr. Lex Fridman
TechnologyArtificial Intelligence (AI)Machine Learning (ML)Deep LearningNeural Networks
NewEconomicThinkingJun 16, 2021The Human Advantage: Framing, Mental Models, and AI in a Complex World
PhilosophyMental Models (Frames)Cognitive PsychologyBig DataMachine Learning
DocSnipesMay 25, 2021Positive Parenting and Behavior Modification: Nurturing Secure Attachment and Developmental Growth
PsychologySecure attachmentCRAVES mnemonicBETA practiceHPA axis down-regulation
lexfridmanJan 18, 2021Max Tegmark on AI, Physics, and the Quest for Intelligible Intelligence
ScienceArtificial General Intelligence (AGI)Existential risks of AIIntelligible intelligenceBlack box AI systems
lexfridmanDec 13, 2020Michael Littman on Reinforcement Learning, AI's Future, and Dispelling Superintelligence Fears
TechnologyReinforcement LearningArtificial Intelligence (AI)Machine LearningArtificial General Intelligence (AGI)
lexfridmanJul 14, 2020Bridging the Intelligence Gap: Sergey Levine on Robotics, Deep Learning, and the Nature of Intelligence
ScienceDeep LearningReinforcement LearningRoboticsComputer Vision
lexfridmanJul 3, 2020The Unified Science of Mind and Brain: Bridging Psychology, Neuroscience, and AI with Matt Botvinick
NeuroscienceCognitive psychologyComputational neuroscienceArtificial intelligenceDeep learning
lexfridmanJun 26, 2020Sophia, AGI, and the Ethics of Anthropomorphic Robotics: A Conversation with Ben Goertzel
TechnologyArtificial General Intelligence (AGI)Computational CompassionSingularityHardware Ecosystem
lexfridmanMay 10, 2020The Unification of AI Modalities: Vision, Language, and Reinforcement Learning, and the Elusive Definition of 'Hardness' in AI
TechnologyDeep LearningComputer VisionNatural Language Processing (NLP)Reinforcement Learning (RL)
lexfridmanMay 9, 2020Ilya Sutskever on Building AGI: Self-Play, Simulation, Consciousness, and Human Control
TechnologyArtificial General Intelligence (AGI)Deep LearningSelf-playReinforcement Learning (RL)
lexfridmanMay 8, 2020Ilya Sutskever on the Deep Learning Revolution, AI's Unity, and the Future of Intelligence
TechnologyDeep Learning RevolutionNeural NetworksBackpropagationHessian-free optimizer
lexfridmanApr 3, 2020David Silver on AlphaGo, AlphaZero, and the Reinforcement Learning Path to Artificial Intelligence
TechnologyReinforcement Learning (RL)Deep Reinforcement LearningAlphaGoAlphaZero
lexfridmanFeb 26, 2020Marcus Hutter on AIXI, Kolmogorov Complexity, and the Computable Universe
TechnologyAIXI modelArtificial General Intelligence (AGI)Kolmogorov complexitySolomonoff induction
lexfridmanFeb 21, 2020Andrew Ng's Expert Advice on Getting Started and Building a Career in Deep Learning and AI
TechnologyDeep LearningMachine LearningNeural NetworksGradient Descent
lexfridmanJan 10, 2020Deep Learning: State of the Art in 2020, Key Advancements, Limitations, and Future Research
TechnologyDeep LearningNeural NetworksPerceptron (single-layer, multi-layer)Backpropagation
lexfridmanOct 9, 2019Machine Learning at Spotify: The Power of User-Generated Playlists for Personalized Music Discovery
TechnologyReinforcement learningState spaceProgramming language (playlists)Meta-programming language
lexfridmanSep 13, 2019George Hotz on Winning, Reinforcement Learning, and Discovering Life's Universal Reward Function
TechnologyReinforcement learningIntelligent agentCompressive model of the worldMaximally compressive
lexfridmanSep 5, 2019George Hotz's Three Problems of Autonomous Driving: Static, Dynamic, and Counterfactual AI Challenges
TechnologyStatic driving problemDynamic driving problemCounterfactual driving problemMapping and localization
DocSnipesAug 21, 2019Stages and Theories of Treatment for the Addiction Counselor & NCMHCE Exam Review
PsychologyStages of Treatment (Initial, Middle, Late, Termination)Cognitive Behavioral Therapy (CBT)BehaviorismHumanistic Therapy
lexfridmanJun 3, 2019TensorFlow's Evolution: From Google Brain's Inception to a Global Open-Source ML Ecosystem
TechnologyDeep LearningTensorFlowOpen SourceGoogle Brain
lexfridmanApr 18, 2019Ian Goodfellow on Generative Adversarial Networks, Deep Learning Limitations, and the Future of AI Cognition
TechnologyGenerative Adversarial Networks (GANs)Deep LearningRepresentation LearningMachine Learning
lexfridmanApr 3, 2019Greg Brockman on OpenAI, AGI, and Shaping the Future of Intelligence
TechnologyArtificial General Intelligence (AGI)Deep LearningReinforcement LearningValue Alignment
lexfridmanMar 12, 2019Leslie Kaelbling on Reinforcement Learning, Planning, Robotics, and the Philosophy of AI
TechnologyReinforcement LearningPlanning (AI)Robot NavigationArtificial Intelligence (AI)
DocSnipesFeb 23, 2019Person-Centered Theory: Understanding Personality, Behavior, and Therapeutic Application
PsychologyPerson-Centered TheorySelf-conceptConditions of WorthEmpathy
lexfridmanFeb 12, 2019Taming the Long Tail: Waymo's Machine Learning Approach to Autonomous Driving Challenges
TechnologyAutonomous drivingMachine learningDeep learningPerception (AV)
lexfridmanJan 19, 2019Tomaso Poggio on Brains, Minds, Machines, and the Nature of Intelligence
ScienceTheory of RelativityGedanken experimentArtificial General Intelligence (AGI)Biological neural networks
DocSnipesAug 23, 2018Behavior Modification Basics: Understanding and Applying Principles for Recovery and Well-being
PsychologyBehavior ModificationStimulus-ResponseReinforcementPunishment
DocSnipesJun 28, 2018Child and Adolescent Development: Key Theories, Practical Strategies, and Environmental Influences
PsychologyBiopsychosocial perspectivePhysiological responsesBehavioral urgesEmotional vocabulary
lexfridmanApr 25, 2018Ilya Sutskever on Deep Learning Foundations, Meta-Learning, and Self-Play for AGI
TechnologyDeep Learning PrinciplesGeneralization Theory (shortest program)Computational Intractability (program search)Backpropagation (circuit search)
lexfridmanJan 25, 2018Deep Reinforcement Learning: From Raw Data to Reasoning and Real-World Action
TechnologyDeep Reinforcement Learning (DRL)Artificial Intelligence (AI) stackEnd-to-end learningSupervised Learning
DocSnipesOct 13, 2017Must-Know Models and Theories of Addiction and Mental Distress for Counselors
PsychologyBeneficenceNon-malfeasancePoly-addictionNeurochemicals
DocSnipesMay 19, 2017Understanding and Managing Triggers and Cravings in Mental Health and Addiction Recovery
PsychologyTrigger (physical stimulus)Trigger (cognitive stimulus)Cognitive distortionsDopamine
lexfridmanSep 27, 2016Deep Reinforcement Learning: Core Methods, Applications, and Policy Gradients
TechnologyDeep Reinforcement Learning (DRL)Reinforcement Learning (RL)Neural NetworksFunction Approximators
DocSnipesSep 27, 2013The Power of Clinical History: Applying Behavioral Principles to Goal Setting and Counseling
PsychologyBehavior modificationTreatment planningMotivationReinforcement (positive)
DocSnipesSep 27, 2013The Power of Clinical History: Applying Behavioral Principles to Goal Setting and Counseling
PsychologyBehavior modificationTreatment planningMotivationReinforcement (positive)
lexfridmanThe Business and Philosophy of Machine Learning: Promise, Limitations, and the Quest for General Intelligence
TechnologyMachine LearningArtificial Intelligence (AI)Deep LearningSupervised Learning
lexfridmanAI's Superhuman Strategy: Nash Equilibrium in Poker and Diplomacy with Noam Brown
TechnologyGame TheoryNash EquilibriumSelf-playCounterfactual Regret Minimization (CFR)
lexfridmanAndrew Ng on Deep Learning, Education, Automation, and the Future of AI Development
TechnologyMachine LearningDeep LearningArtificial Intelligence (AI)Automation
lexfridmanAlphaZero, Self-Play, and the Path to General AI
TechnologySelf-playReinforcement LearningAlphaGo ZeroAlphaZero
hubermanlabThe Science & Treatment of Obsessive-Compulsive Disorder (OCD) and its Distinction from Obsessive-Compulsive Personality Disorder (OCPD)
PsychologyObsessive-Compulsive Disorder (OCD)Obsessive-Compulsive Personality Disorder (OCPD)ObsessionsCompulsions
lexfridmanProgramming Meme Review with George Hotz: Automation, AI, and the Culture of Code
TechnologyProgramming memesAutomation paradoxSelf-driving carsKnowledge outsourcing
lexfridmanOriol Vinyals on Generalist AI Agents, AGI, and the Future of Deep Learning
TechnologyNeural NetworksDeep LearningArtificial General Intelligence (AGI)Reinforcement Learning
lexfridmanPractical Deep Learning Strategies: Scale, End-to-End Architectures, and Evolving Bias-Variance Trade-offs
TechnologyDeep Learning ProgressScale (Data & Computation)Traditional Learning AlgorithmsNeural Networks
lexfridmanOriol Vinyals on AlphaStar, Deep Reinforcement Learning, and the Evolution of AI in Real-Time Strategy Games
TechnologyDeep LearningReinforcement Learning (RL)AlphaStarStarCraft (Real-Time Strategy game)
psychacksThe Power of Shaping: Never Punish People for Doing What You Want
PsychologyShapingSuccessive approximationsReinforcement contingenciesPositive reinforcement
lexfridmanMIT 6.S091 Introduction to Deep Reinforcement Learning: Fundamentals, Challenges, and Applications
TechnologyDeep Reinforcement Learning (Deep RL)Deep Neural NetworksSequential DecisionsTrial and Error
psychacksAs Little as Possible: Why Men Need to Do the Bare Minimum in Courtship
RelationshipsBare minimum strategyCourtship processRelationship standardsAcclimation effect
psychacksHow to Cultivate a Fulfilling Sexual Relationship by Navigating Libido Differences and Establishing Early Boundaries
RelationshipsLibido compatibilitySexual driveSexual moodReinforcement protocols
lexfridmanMIT 6.S094: Introduction to Deep Learning, Neural Networks, and Self-Driving Cars
TechnologyDeep LearningDeep Neural NetworksSelf-Driving CarsDeep Reinforcement Learning
psychacksSetting Effective Boundaries: Leveraging Personal Agency and Consequences
PsychologyEffective boundary settingIneffective boundary settingReactance (psychological phenomenon)Personal autonomy
psychacksNavigating the Extinction Burst: Why Changing Dysfunctional Relationship Dynamics Gets Worse Before It Gets Better
PsychologyExtinction burstBehavioral psychologyOperant conditioningReinforcement (habitual)
lexfridmanMIT AGI: Engineering Intelligence - Bridging Theory, Practice, and Societal Impact
TechnologyArtificial General Intelligence (AGI)Deep LearningReinforcement LearningComputational Cognitive Science
lexfridmanReverse Engineering Human Intelligence for Advanced AI and AGI: A Cognitive Science Approach
ScienceArtificial General Intelligence (AGI)AI TechnologiesDeep LearningPattern Recognition
lexfridmanMIT 6.S094: Deep Learning for Self-Driving Cars - Foundations, Challenges, and Human-Centered AI
TechnologyDeep LearningSelf-Driving CarsArtificial Intelligence (AI)Reinforcement Learning
Knowledge Graph
Related concepts — line thickness indicates connection strength. Click any node to explore.
Want to explore how this connects to concepts you choose? Try Nexus →