Evaluating and improving Replit Agent at scale

19 mag 202610:30 - 10:45
Tech Stage

Description

Building AI agents that work reliably at scale is one of the hardest challenges in applied AI. In this talk, Michele explores how Replit approaches evaluation and continuous improvement of Replit Agent: the AI-powered coding assistant used by millions of builders worldwide. From designing meaningful evals to closing the feedback loop between real-world usage and model improvements, Michele shares the frameworks, lessons learned, and open questions from operating an AI agent at production scale.

REPLIT

Replit è una piattaforma di sviluppo agentico che permette a chiunque di creare applicazioni tramite linguaggio naturale. Con milioni di utenti e oltre 500.000 business user, rende lo sviluppo...