Optimizing Agent Behavior in Production with Gideon Mendels

LLM -powered systems continue to move steadily into production, but this process is presenting teams with challenges that traditional software practices don’t commonly encounter. Models and agents are non-deterministic systems, which makes it difficult to test changes, reason about failures, and confidently ship updates. This has created the need for new evaluation tooling designed specifically around the properties of LLMs.

Comet is a platform with Roots and MLOps, to the rapidly evolving world of agent-based systems by treating prompts, tools, and workflows as optimizable components that can be evaluated and improved over time.

Gideon Mendels is the co -founder and CEO of Comet. He previously worked at Google on hate speech and deception detection, and he founded GroupWise, which trained and deployed NLP models processing billions of chats. In this episode, Gideon joins Kevin Ball to discuss how agent development sits between software engineering and ML, why eVals are the missing foundation for most AI teams, prompt optimization as a search problem, and the future for continuously improving agents in production.

Full Disclosure: This episode is sponsored by Comet.

lee 6

Kevin Ball or KBall, is the vice president of engineering at Mento and an independent coach for engineers and engineering leaders. He co-founded and served as CTO for two companies, founded the San Diego JavaScript meetup, and organizes the AI inaction discussion group through Latent Space.

Please click here to see the transcript of this episode.

Source link

What's Hot

Amazon plans huge AWS investment to meet AI cloud demand

Apple Spring Event 2026: Date, time, how to watch, and what to expect

Checkmarx Enhances IDE-Native Agentic Application Security in Kiro

Optimizing Agent Behavior in Production with Gideon Mendels

Infostealer Steals OpenClaw AI Agent Configuration Files and Gateway Tokens

Open source maintainers are being targeted by AI agent as part of ‘reputation farming’

How an AI agent is redefining executive workflows at Cemex

Amazon plans huge AWS investment to meet AI cloud demand

Apple Spring Event 2026: Date, time, how to watch, and what to expect

Checkmarx Enhances IDE-Native Agentic Application Security in Kiro

Recent Advances in Lithium Metal Protective Strategies with Stable Interface

Don't Miss!

Amazon plans huge AWS investment to meet AI cloud demand

Apple Spring Event 2026: Date, time, how to watch, and what to expect

Subscribe to Updates

What's Hot

Optimizing Agent Behavior in Production with Gideon Mendels

Related Posts

Subscribe to Updates