PostHole
Compose Login
You are browsing us.zone2 in read-only mode. Log in to participate.
rss-bridge 2026-02-17T10:00:15+00:00

Optimizing Agent Behavior in Production with Gideon Mendels

LLM -powered systems continue to move steadily into production, but this process is presenting teams with challenges that traditional software practices don’t commonly encounter. Models and agents are non-deterministic systems, which makes it difficult to test changes, reason about failures, and confidently ship updates. This has created the need for new evaluation tooling designed specifically
The post Optimizing Agent Behavior in Production with Gideon Mendels appeared first on Software Engineering Daily.


**
**
**

Optimizing Agent Behavior in Production with Gideon Mendels

By SEDaily

** Podcast
Tuesday, February 17 2026

Podcast: Play in new window | Download

Subscribe: RSS

LLM -powered systems continue to move steadily into production, but this process is presenting teams with challenges that traditional software practices don’t commonly encounter. Models and agents are non-deterministic systems, which makes it difficult to test changes, reason about failures, and confidently ship updates. This has created the need for new evaluation tooling designed specifically around the properties of LLMs.

Comet is a platform with Roots and MLOps, to the rapidly evolving world of agent-based systems by treating prompts, tools, and workflows as optimizable components that can be evaluated and improved over time.

Gideon Mendels is the co -founder and CEO of Comet. He previously worked at Google on hate speech and deception detection, and he founded GroupWise, which trained and deployed NLP models processing billions of chats. In this episode, Gideon joins Kevin Ball to discuss how agent development sits between software engineering and ML, why eVals are the missing foundation for most AI teams, prompt optimization as a search problem, and the future for continuously improving agents in production.

Kevin Ball or KBall, is the vice president of engineering at Mento and an independent coach for engineers and engineering leaders. He co-founded and served as CTO for two companies, founded the San Diego JavaScript meetup, and organizes the AI inaction discussion group through Latent Space.

Please click here to see the transcript of this episode.

Sponsorship inquiries: sponsor@softwareengineeringdaily.com

SEDaily

****POPULAR****

Software Daily

Subscribe to Software Daily, a curated newsletter featuring the best and newest from the software engineering community.

Exclusive Articles

VMware Tanzu GemFire and Next-Generation Real-Time Application Development
Uber’s LedgerStore and its Trillions of Indexes with Kaushik Devarajaiah
GraphQL vs. REST: What Are They, and Which Is Better for You?

Cloud Engineering

CodeRabbit and RAG for Code Review with Harjot Gill
Building Chess.com with Jay Severson
Mastodon with Eugen Rochko

Business and Philosophy

Startup Investing with George Mathew
KubeCon Special: Docker with Justin Cormack
Software Architecture with Josh Prismon

Greatest Hits

Hardening C++ with Bjarne Stroustrup
Surviving ChatGPT with Christian Hubicki
Special Episode with George Hotz

Hackers

Making React 70% faster with Aiden Bai of Million.js
Cross-functional Incident Management with Ashley Sawatsky and Niall Murphy
SDKs for your API with Sagar Batchu

Data

Hyperscaling SQL with Sam Lambert
Spring AI and Java in 2024
Iceberg at Netflix and Beyond with Ryan Blue


*Original source*

Reply