
20 days of compute vs 7 hours: rethinking what state-of-the-art means — Bertrand Charpentier, Pruna AI
Bertrand Charpentier
Vision and video · Industry applications · Evals
AI Engineer topic
272 AI Engineer conference talks about Evals.

Bertrand Charpentier
Vision and video · Industry applications · Evals

John Dickerson
Safety and governance · Evals · Reasoning and models

Harrison Chase
Evals · Coding and developer tools · Agent engineering

Hamed Firooz · Maziar Sanjabi
RAG, context, and search · Evals · Infrastructure and deployment

Justin Muller
Reasoning and models · Evals · Observability and reliability

Logan Kilpatrick
APIs, MCP, and protocols · Evals · Reasoning and models

Chintan Parikh · Weiyi Wang
Evals · Other / unclassified · Infrastructure and deployment

Sara Hooker
Infrastructure and deployment · Reasoning and models · Industry applications

Ari Heljakka
Evals · RAG, context, and search · APIs, MCP, and protocols

Will Hang · Cathy Zhou
Data and model adaptation · Agent engineering · Coding and developer tools

Nicholas Kang · Michael Aaron
Evals · Agent engineering · Creative and generative media

Cedric Vidal
Safety and governance · Evals · Coding and developer tools

Uday Kiran Medisetty · Adam Huda
Coding and developer tools · Safety and governance · Evals

Jesse Hu
Enterprise · RAG, context, and search · Data and model adaptation

Dan Fu · Olive Song
Evals · Architecture · Agent engineering

Alfonso Graziano
Architecture · Agent engineering · Evals

Mike Spitz
Architecture · Evals · Enterprise

Sachin Gupta
Agent engineering · Evals · RAG, context, and search

Ian Butler · Nick Gregory
Coding and developer tools · Infrastructure and deployment · Evals

Anita Kirkovska
RAG, context, and search · Coding and developer tools · Observability and reliability

Nathaniel Whittemore (NLW)
Leadership · Infrastructure and deployment · Evals

Varsha Shah
Finance · Enterprise · RAG, context, and search

Charles Frye
Enterprise · Evals · RAG, context, and search

Natalie Serrino
Infrastructure and deployment · Evals · Other / unclassified

Nagkumar Arkalgud · Keiji Kanazawa
RAG, context, and search · Safety and governance · Evals

Apoorva Joshi
Safety and governance · Evals · Architecture

Brendan Rappazzo
RAG, context, and search · Finance · Agent engineering

Alexander Bricken · Joe Bayley
Coding and developer tools · Infrastructure and deployment · Evals

Denys Linkov
Agent engineering · Coding and developer tools · Healthcare

Alex Duffy
Evals

Ali Khial
Evals · RAG, context, and search · Coding and developer tools

Niklas Nielsen
Evals · Coding and developer tools · Other / unclassified

Marlene Mhangami
Coding and developer tools · Developer workflows and testing · APIs, MCP, and protocols

Parth Asawa
RAG, context, and search · Evals · Safety and governance

Paul Henry
RAG, context, and search · Evals · Infrastructure and deployment

Aparna Dhinakaran
Evals · Reasoning and models · RAG, context, and search

Paul Klein IV
Agent engineering · Coding and developer tools · Evals

Samuel Denton
Data and model adaptation · Enterprise · Safety and governance

SallyAnn DeLucia · Fuad Ali
Observability and reliability · RAG, context, and search · Evals

Nick Ung · Akshay Sharma
Evals · Agent engineering · Infrastructure and deployment

Louis-François Bouchard · Paul Iusztin · Samridhi Vaid
Creative and generative media · Agent engineering · Coding and developer tools

Anant Dole · Asbjørn Steinskog
Agent engineering · Other / unclassified · Creative and generative media

Ben Hylak · Sid Bendre
Enterprise · Evals · Reasoning and models

Michael Albada
Observability and reliability · Architecture · Evals

Eugene Yan
RAG, context, and search · Evals · Enterprise

Soumya Gupta · Jai Chopra
Observability and reliability · Evals · Safety and governance

Cedric Vidal
Evals · Coding and developer tools · Reasoning and models

Shaan Desai
Evals · RAG, context, and search · Architecture