Skip to content
all episodes
Research Insights Made Simple · episode 11

Measuring AI Code Assistants and Agents

Measuring AI Code Assistants and Agents

July 8, 202533:36

Episode participants

What we discussed

An episode on measuring AI code assistants and agents: what counts as outcome and where benchmark-style evaluation becomes misleading.

The discussion covers tasks, evals, change quality, security, review load and how agentic workflows change development metrics.

The focus is moving from local coding speed to the impact on the whole SDLC.

AI in SDLCDeveloper productivityResearch methodology