# Events & Webinars

Learn from experts, watch demos, and stay updated on the latest in AI quality and safety.

2025-06-03 [**Will Agent Evaluation via MCP Stabilize Agent Frameworks?**](https://scorable.ai/post/will-agent-evaluation-via-mcp-stabilize-agent-frameworks)

Discover how exposing complex AI Evaluation frameworks to agents via MCP (Model Context Protocol) allows for a new paradigm of controllable self-improvement.

[Watch recording →](https://scorable.ai/post/will-agent-evaluation-via-mcp-stabilize-agent-frameworks)

2025-03-13 [**Empowering AI with Cost-Effective LLM Judges**](https://scorable.ai/post/empowering-ai-with-cost-effective-llm-judges)

Explore the dual themes of agent evaluations and EvalOps in this comprehensive technical session on cost-effective LLM judges.

[Watch recording →](https://scorable.ai/post/empowering-ai-with-cost-effective-llm-judges)

2025-02-19 [**Agent Evals: Finally, With The Map**](https://scorable.ai/post/agent-evals-finally-with-the-map)

A comprehensive look at agent evaluation frameworks and methodologies, delivered at the AI Engineer Summit: Agents at Work!

[Watch recording →](https://scorable.ai/post/agent-evals-finally-with-the-map)

2025-02-06 [**10 Critical LLM Blunders - Detect and Fix with LLM Judges**](https://scorable.ai/post/10-critical-llm-blunders)

Dive deep into the intricacies of LLM-based applications and learn to detect, block, and remedy the most common yet critical errors that undermine reliability.

[Watch recording →](https://scorable.ai/post/10-critical-llm-blunders)

2024-12-04 [**TOP-10 Misconceptions about LLM Judges in Production**](https://scorable.ai/post/top-10-misconceptions-about-llm-judge)

Debunking common myths and misconceptions about implementing LLM judges in production environments based on real-world experience.

[Watch recording →](https://scorable.ai/post/top-10-misconceptions-about-llm-judge)

2024-11-07 [**EvalOps - Mastering The Game of LLM Judges**](https://scorable.ai/post/evalops-101)

A comprehensive keynote on operational excellence in LLM evaluation and judgment systems.

[Watch recording →](https://scorable.ai/post/evalops-101)

2024-09-18 [**Building Your Optimal LLM Evaluation Stack**](https://scorable.ai/post/building-your-optimal-llm-evaluation-stack)

Learn how to create a robust framework for evaluating and optimizing large language models, covering best practices, tools, and strategies for production reliability.

[Watch recording →](https://scorable.ai/post/building-your-optimal-llm-evaluation-stack)
