all episodes
Research Insights Made Simple · episode 11

Measuring AI Code Assistants and Agents

Measuring AI Code Assistants and Agents

July 8, 202533:36

Episode participants

  • Alexander Polomodov

    host

  • Евгений Сергеев

    guest · Engineering Director · Flo Health

    Евгений Сергеев — Engineering Director в Flo Health.

What we discussed

An episode on measuring AI code assistants and agents: what counts as outcome and where benchmark-style evaluation becomes misleading.

The discussion covers tasks, evals, change quality, security, review load and how agentic workflows change development metrics.

The focus is moving from local coding speed to the impact on the whole SDLC.

AI in SDLCDeveloper productivityResearch methodology