Haize Labs with Leonard Tang - Weaviate Podcast #121!

Weaviate vector database · Intermediate ·🛡️ AI Safety & Ethics ·12mo ago

Skills: AI Alignment Basics90%

How do you ensure your AI systems actually do what you expect them to do? Leonard Tang takes us deep into the revolutionary world of AI evaluation with concrete techniques you can apply today. Learn how Haize Labs is transforming AI testing through "scaling judge-time compute" - stacking weaker models to effectively evaluate stronger ones. Leonard unpacks the game-changing Verdict library that outperforms frontier models by 10-20% while dramatically reducing costs. Discover practical insights on creating contrastive evaluation sets that extract maximum signal from human feedback, implementing debate-based judging systems, and building custom reward models that align with enterprise needs. The conversation reveals powerful nuggets like using randomized agent debates to achieve consensus and lightweight guardrail models that run alongside inference. Whether you're developing AI applications or simply fascinated by how we'll ensure increasingly powerful AI systems perform as expected, this episode delivers immediate value with techniques you can implement right away, philosophical perspectives on AI safety, and a glimpse into the future of evaluation that will fundamentally shape how AI evolves. Learn more about Haize Labs! - https://www.haizelabs.com/ Check out Verdict on GitHub - https://github.com/haizelabs/verdict Chapters 0:00 Weaviate Podcast #121! 0:46 Welcome Leonard! 1:16 Founding Haize Labs 8:31 UX for Evals 17:01 Scaling Judge-Time Compute with Verdict 23:26 Debate Judges 26:06 Compute Scaling 28:50 Declarative Judge Pipelines 31:13 Custom Reward Models 37:20 Reasoning in Reward Models 39:20 Mechanistic Interpretability 45:30 Guardrails in Inference Pipelines 47:35 Can we control Superintelligence? 52:08 Exciting Directions for AI

Watch on YouTube ↗ (saves to browser)

Sign in to unlock AI tutor explanation · ⚡30

More on: AI Alignment Basics

View skill →

Interpretable machine learning applications: Part 5

Interpretable machine learning applications: Part 5

GenAI news from Weights & Biases CEO, Lukas Biewald

GenAI news from Weights & Biases CEO, Lukas Biewald

Weights & Biases

Responsible AI Winners, 2020 PyTorch Summer Hackathon

Responsible AI Winners, 2020 PyTorch Summer Hackathon

Near Real-Time Analytics to GenAI Centralized Observability | Amazon Web Services

Near Real-Time Analytics to GenAI Centralized Observability | Amazon Web Services

Amazon Web Services

Kiro Hooks | Event-Driven Automation for Your IDE | Amazon Web Services

Kiro Hooks | Event-Driven Automation for Your IDE | Amazon Web Services

Amazon Web Services

Get Started with Raven AGI

Get Started with Raven AGI

Related AI Lessons

Behind the Scenes Hardening Firefox with Claude Mythos Preview

Learn how Mozilla used Claude Mythos to identify and fix hundreds of vulnerabilities in Firefox, improving browser security

Simon Willison's Blog

AI Alignment Might Be Optimizing the Wrong Objective

AI alignment might be optimizing the wrong objective, highlighting the need to redefine what alignment means and how it's achieved

AI Alignment Might Be Optimizing the Wrong Objective

AI alignment might be optimizing the wrong objective, highlighting the need to redefine what alignment means and how it's achieved

Medium · Machine Learning

Cognitive Surrender: how much thinking should leaders outsource to AI?

Learn how leaders can effectively balance AI-driven insights with human judgment to avoid cognitive surrender

Medium · Data Science

Chapters (14)

Weaviate Podcast #121!

0:46 Welcome Leonard!

1:16 Founding Haize Labs

8:31 UX for Evals

17:01 Scaling Judge-Time Compute with Verdict

23:26 Debate Judges

26:06 Compute Scaling

28:50 Declarative Judge Pipelines

31:13 Custom Reward Models

37:20 Reasoning in Reward Models

39:20 Mechanistic Interpretability

45:30 Guardrails in Inference Pipelines

47:35 Can we control Superintelligence?

52:08 Exciting Directions for AI

Why you can’t love all animals and still eat meat