
Best Practices for Evaluating Large Language Model Applications with llmeval: Niklas Nielsen
Niklas Nielsen
Evals · Coding and developer tools · Other / unclassified
1,011 published talks.

Niklas Nielsen
Evals · Coding and developer tools · Other / unclassified

Arjun Chintapalli · Bhavani Kalisetty
Agent engineering · Data and model adaptation · RAG, context, and search

Marlene Mhangami
Coding and developer tools · Developer workflows and testing · APIs, MCP, and protocols

Filip Kozera
Agent engineering · Coding and developer tools · RAG, context, and search

Parth Asawa
RAG, context, and search · Evals · Safety and governance

Grace Isford
Agent engineering · Vision and video · Infrastructure and deployment

Rajiv Chandegra
Agent engineering · Industry applications

Josh Albrecht
Coding and developer tools · Developer workflows and testing

Hervé Bredin
Speech and audio

Stephen Batifol
Creative and generative media · Robotics and world models

Siddharth Ahuja
APIs, MCP, and protocols · Vision and video · RAG, context, and search

Eric Simons
Enterprise · Other / unclassified

Łukasz Gandecki
Coding and developer tools · Enterprise · RAG, context, and search

Paul Henry
RAG, context, and search · Evals · Infrastructure and deployment

Angus J. McLean
Agent engineering · Creative and generative media · RAG, context, and search

Aparna Dhinakaran
Evals · Reasoning and models · RAG, context, and search

Sunny Madra
Infrastructure and deployment · Safety and governance · Enterprise

Greg Benson
Agent engineering · Reasoning and models · Coding and developer tools

Paul Klein IV
Agent engineering · Coding and developer tools · Evals

Samuel Denton
Data and model adaptation · Enterprise · Safety and governance

SallyAnn DeLucia · Fuad Ali
Observability and reliability · RAG, context, and search · Evals

Angel Ortmann Lee
Agent engineering · Safety and governance · Other / unclassified

Paige Bailey
Enterprise · Vision and video · Speech and audio

Paige Bailey · Guillaume Vernade · Ian Ballantyne
Creative and generative media · Infrastructure and deployment · RAG, context, and search

Eliza Cabrera · Jeremy Silva
Enterprise · Reasoning and models · RAG, context, and search

Varun Badrinath Krishna · Petro Junior Milan · Rachelle Mattern
RAG, context, and search · Safety and governance · Reasoning and models

Nick Ung · Akshay Sharma
Evals · Agent engineering · Infrastructure and deployment

Cedric Vidal · David Smith · Miguel Martinez
Safety and governance · Other / unclassified · Reasoning and models

Shawn Chan
Finance · Leadership · Observability and reliability

Angie Jones
Infrastructure and deployment · Agent engineering · Coding and developer tools

Dr. Sajjan Kanukolanu
Architecture · Safety and governance · Enterprise

Raj Navakoti
RAG, context, and search · Coding and developer tools · Observability and reliability

Louis-François Bouchard · Paul Iusztin · Samridhi Vaid
Creative and generative media · Agent engineering · Coding and developer tools

Max Brodeur-Urbas
Reasoning and models · Infrastructure and deployment · Leadership

Anant Dole · Asbjørn Steinskog
Agent engineering · Other / unclassified · Creative and generative media

Will Bryk
RAG, context, and search · Architecture