cost
65 talks
PodcastAutonomous Agents at Work: From OpenClaw Hype to Enterprise Reality
PodcastAgents & the $40M Bet on Multiplayer AI
What's Special About Meta's Multi-Agent Systems
PodcastHow We Cut LLM Latency 70% With TensorRT in Production
The Coding Agent Multiverse of Madness
The Shadow AI Problem Nobody's Talking About
How to Optimize AI Agents in Production
What It Takes to Run Multi-Agent Systems
The Cost of AI: FinOps Strategies for Intelligent Agents
AI Needs Memory: Here's How It Works
Advancing the Cost-Quality Frontier in Agentic AI
Cutting Costs with Artificial Intelligence
Reading groupSmall Language Models are the Future of Agentic AI
The Future of Compute: How AI Agents Are Reshaping Infrastructure
Smart Agents Start with Smart LLM Choices
How Agents Changed Vibe Coding Forever
Testing AI Intelligence: The Benchmarking Battle
PodcastStreaming Ecosystem Complexities and Cost Management
PodcastThe Agent Landscape - Lessons Learned Putting Agents Into Production
PodcastReal World AI Agent Stories
We're Using AI Agents at Work (and it's amazing)
PodcastAI-Driven Code: Navigating Due Diligence & Transparency in MLOps
LLMs to agents: The Beauty & Perils of Investing in GenAI
How to Actually Use Cost Effective AI in Your Business
AI-Powered Data Unification for Data Platforms
How To Cut Your Data Infrastructure Costs in Half
PodcastAWS Trainium and Inferentia
Navigating the Emerging LLMOps Stack
No GPU Before PMF
Productionizing AI: How to Think From the End
LLMOps and GenAI at Enterprise Scale - Challenges and Opportunities
Streamlining Model Deployment
Reliable Hallucination Detection in Large Language Models
Anatomy of a Software 3.0 Company
Building RAG-based LLM Applications for Production
Current State of LLMs in Production
Efficient Serving of LLMs for Experimentation and Production with Fireworks.ai
Amplifying Impact with Generative AI: Insights from 10,000 Colleagues
Exploring the Latency/Throughput & Cost Space for LLM Inference
AI in Education Fireside Chat
Finetuning Open-Source LLMs
Fireside Chat with LLM Startups
What Drives GenAI Development in the Next 3 Years
PodcastTecton Round-table // Get your ML Application Into Production
PodcastFrugalGPT: Better Quality and Lower Cost for LLM Applications
Preemption Chaos and Optimizing Server Startup
LLM XGBoost: Can a Fine-Tuned LLM Beat XGBoost on Tabular Data?
The Confidence Checklist for LLMs in Production
Making LLM Inference Affordable
Building Products
PodcastTreating Prompt Engineering More Like Code
It Worked When I Prompted It
Understanding the LLM Economics
Taking ImgFlip's 'This Meme Does Not Exist' to the Next Level with a LLM
The Emerging Toolkit for Reliable, High-quality LLM Applications
PodcastFrom Arduinos to LLMs: Exploring the Spectrum of ML
PodcastThe Long Tail of ML Deployment
PodcastWhy is MLOps Hard in an Enterprise?
Using LLMs to Punch Above Your Weight!
Efficiently Scaling and Deploying LLMs
Cost Optimization and Performance
PodcastBringing DevOps Agility to ML
PodcastReal-time Model Inference in a Video Streaming Environment
PodcastBuilding ML/Data Platform on Top of Kubernetes
MeetupLaw of Diminishing Returns for Running AI Proof-of-Concepts