reinforcement-learning

13 talks

Coding Agents Are Secretly General AgentsJay Hack, ClickUp · 1:12:03 · Jul 2026 · 243 viewsInside OpenAI's AI Agent Collaboration System · 19:22 · Jan 2026 · 201 viewsHow Reinforcement Learning Can Improve Your AgentPatrick Barker · 12:25 · Aug 2025 · 205 views · Agents in Production 2025PodcastTricks to Fine TuningPrithviraj Ammanabrolu, Databricks · 54:02 · May 2025 · 290 views · MLOps PodcastPodcastFrom Shiny to Strategic: The Maturation of AI Across IndustriesDavid Cox, RethinkFirst; Institute of Applied Behavioral Science · 40:51 · Apr 2025 · 103 views · MLOps PodcastPodcastBeyond the Matrix: AI and the Future of Human CreativityFausto Albers, AI Builders Club · 55:09 · Mar 2025 · 202 views · MLOps PodcastReading groupDeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement LearningAdam Becker, MLOps Community & Nehil Jain, Stealth AI Startup & Matt Squire, Fuzzy Labs & Sophia Skowronski, Breckinridge Capital Advisors · 1:00:25 · Mar 2025 · 551 views · MLOps Reading GroupWhy Agents Are Stupid & What We Can Do About ItDan Jeffries, Kentauros AI · 31:58 · Dec 2024 · 1,150 viewsGoal Oriented Retrieval AgentsZoe Weil, Faber Labs · 25:09 · Dec 2024 · 750 views · Agents in Production 2024PodcastPyTorch for Control Systems and Decision MakingVincent Moens, Meta · 55:26 · Dec 2024 · 564 views · MLOps PodcastExplaining ChatGPT to Anyone in 10 MinutesCameron Wolfe, Rebuy · 11:16 · May 2024 · 380 views · AI in Production 2024RLHF Data Collection in PracticeAndrew Mauboussin, Surge AI · 12:10 · Aug 2023 · 684 views · LLMs in Production 2023Building and Curating Datasets for RLHF and LLM Fine-tuningDaniel Vila Suero, Argilla · 58:51 · Jul 2023 · 3,363 views · LLMs in Production 2023