human-in-the-loop
82 talks
MCPs for Observability Stacks
2026 The Year of Agent Orchestration
How to Make a Coding Agent a General Purpose Agent
Everything We Got Wrong About Research-Plan-Implement
Stop Building AI Like Traditional Software
Structured Dissent Patterns for Agentic Production Reliability
PodcastEnterprise AI Operations: The Missing Piece
Fine-Tuned Models Are Getting Out of Hand
Evals Aren't Useful? Really?
Underwriting Assist: A Multi-Agent System
Why You Should Care About Observability in LLM Workflows
Designing AI Agents for the Complex Realities of Healthcare
Catastrophic agent failure and how to avoid it
PodcastThe Era of AI Agents in Marketing
If There's Free Compute, There's Abuse: Fighting Fraud with Lightweight LLM Agents
Fast, Trustworthy, Reliable Voice Agents: MLOps That Blend LLM Annotation with Human QA
Too much lock-in for too little gain: agent frameworks are a dead-end
Iterating on Your AI Evals
How AI Will Transform The Energy Sector
I Built A Trustworthy Voice Assistant
PodcastKnowledge is Eventually Consistent
How Product Metrics Become LLM Evaluations
MCP is not going to change everything (yet)
Building an AI agent with LangGraph, step by step tutorial
PodcastMaking AI Reliable is the Greatest Challenge of the 2020s
PodcastAI Data Engineers: Data Engineering After AI
PodcastHow Sama is Improving ML Models to Make AVs Safer
PodcastAI-Powered Product Ideation with Synthetic Consumer Testing
Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations
PodcastLook At Your ****ing Data 👀
PodcastEvolving Workflow Orchestration
Building Reliable Agents
Hundreds of Users Love Our Data Analyst AI Agent
Why Planning is the New Search
The Future of Healthcare: AI is Here
Cleric AI SRE: Towards Self-healing Autonomous Software
PodcastHow Agentic Workflows Will Change Everything
The Next Revolution in AI: LLMs and Beyond
Turn Data Chaos into AI Strategy with Programmatic AI Data Development
Vision and Strategies for Attracting & Driving AI Talents in High Growth
Balancing Speed and Safety
PodcastReliable LLM Products, Fueled by Feedback
Ghostwriter - AI Writing That Learns From You
Data Labeling Best Practices
Graphs and Language
From Research to Production: Fine-Tuning & Aligning LLMs
PodcastLanguage, Graphs, and AI in Industry
PodcastModel Management in a Regulated Environment
AI Squared: Breaking LLMs out of the Chat Application
Assess the Value and Feasibility of LLM Use Cases with a Checklist
From Building Self-driving Cars to Building LLM Applications
LLMs in Production at GetYourGuide
Amplifying Impact with Generative AI: Insights from 10,000 Colleagues
Automating Data Annotation with LLMs
Fireside Chat - The Future of LLMs
PodcastUsing Large Language Models at AngelList
Evolving AI Governance for an LLM World
UX of an LLM User
Incorporating LLMs in High-stake Use Cases
RLHF Data Collection in Practice
PodcastMLOps at the Age of Generative AI
Designing Human in the Loop Experiences for LLMs
Building Reliable AI Agents
Evaluation
Building and Curating Datasets for RLHF and LLM Fine-tuning
LLMs as Intelligent Assistants
Guiding LLMs While Staying in the Driver's Seat
Large Language Models in Production Round-table Conversation
MeetupCreative AI: Using ML to Create Art, Music, and Jokes
PodcastLet's Continue Bundling into the Database
PodcastScaling Similarity Learning at Digits
PodcastLabeled Datasets that Correct Themselves Automatically
PodcastMLOps Critiques
MeetupApplications of Data Science
PodcastThe Future of AI and ML in Process Automation
MeetupData-Centric AI Means Centralizing Training Data
PodcastData Selection for Data-Centric AI: Data Quality Over Quantity
MeetupEngineering MLOps
MeetupProduct Management in Machine Learning
MeetupAgile AI Ethics: Balancing Short Term Value with Long Term Ethical Outcomes
MeetupCreating Beautiful Ambient Music with Google Brain's Music Transformer
MeetupScaling Human-in-the-Loop Machine Learning