AI Engineer knowledge library

Organizations.

Discover the companies, research groups, and institutions represented in AI Engineer conference talks. Affiliations reflect the context of each recorded session.

614 organizations998 talks993 speakers

Microsoft

29 talks · 27 speakers

Accelerate your AI journey with Azure AI model catalog · Agentic Excellence: Mastering Evaluation of AI Agents with Azure AI Evaluation SDK

Google DeepMind

27 talks · 27 speakers

A year of Gemini progress + what comes next · Agentic Evaluations at Scale — For Everybody

Anthropic

23 talks · 26 speakers

Don't Build Agents, Build Skills Instead · Anthropic for VPs of AI

OpenAI

22 talks · 25 speakers

Agent Reinforcement Fine Tuning · Building Effective Voice Agents

Neo4j

17 talks · 11 speakers

Agentic GraphRAG: AI’s Logical Edge · Agentic GraphRAG: Simplifying Retrieval Across Structured & Unstructured Data — Zach Blumenfeld

GitHub

15 talks · 15 speakers

Collaborating with Agents in Your Software Development Workflow - Jon Peck & Christopher Harrison, GitHub · Copilots Everywhere

Arize

14 talks · 10 speakers

Break It 'Til You Make It: Building the Self-Improving Stack for AI Agents · Build a Prompt Learning Loop

NVIDIA

14 talks · 19 speakers

Compression at the Edge · #define AI Engineer

Braintrust

13 talks · 8 speakers

Does GenAI "belong" to data scientists? · Evals 101 — Doug Guthrie, Braintrust

Google

10 talks · 8 speakers

Accelerating AI on Edge — Chintan Parikh and Weiyi Wang, Google DeepMind · Building Agent Interfaces: Lessons from Chrome DevTools (MCP) for Agents

WorkOS

9 talks · 5 speakers

Agents[REDACTED:location_address] Access[REDACTED:location_address] and the Future of Machine Identity · AI Pipelines and Agents in Pure TypeScript with Mastra.ai

Cloudflare

7 talks · 6 speakers

Agents[REDACTED:location_address] Access[REDACTED:location_address] and the Future of Machine Identity · Building Agents (the hard parts!)

Daily

7 talks · 5 speakers

Full Workshop: Realtime Voice AI — Mark Backman, Daily · How to build the world's fastest voice bot

Hugging Face

7 talks · 6 speakers

Compression at the Edge · How I automate my own job at Hugging Face using agents

Amazon

6 talks · 3 speakers

Agent Output Is Not UX: Rendering Layer Your LLM Pipeline Is Missing - Bala Ramdoss, Amazon Lens · Building Blocks for LLM Systems & Products

Amazon Web Services (AWS)

6 talks · 6 speakers

7 Habits of Highly Effective Generative AI Evaluations · Data is Your Differentiator: Building Secure and Tailored AI Systems

AWS

6 talks · 7 speakers

AI Didn’t Kill the Web, It Moved in! — Olivier Leplus (AWS) & Yohan Lasorsa (Microsoft) · Building Agents at Cloud Scale — Antje Barth, AWS

Cursor

6 talks · 5 speakers

Building Cursor Composer · Building your own software factory

Latent.Space

6 talks · 5 speakers

AI Engineering Without Borders · Designing AI-Intensive Applications

MongoDB

6 talks · 5 speakers

AI System Design: From Idea to Production · Architecting Agent Memory: Principles, Patterns, and Best Practices

Vercel

6 talks · 5 speakers

AI SDK v6 · Building durable Agents with Workflow DevKit & AI SDK

AI Engineer

5 talks · 2 speakers

6 Things to Know about AIE World's Fair 2026 · Agents for Everything Else — swyx

Augment Code

5 talks · 5 speakers

Building Self-Coding Agents · Frontier Feud

ElevenLabs

5 talks · 6 speakers

Building Conversational AI Agents - Thor Schaeff, ElevenLabs · Give Your Chat Agent a Voice

HumanLayer

5 talks · 2 speakers

12-Factor Agents: Patterns of reliable LLM applications · No Vibes Allowed: Solving Hard Problems in Complex Codebases

LlamaIndex

5 talks · 2 speakers

Building AI Agents that actually automate Knowledge Work · Effective agent design patterns in production

Modal

5 talks · 3 speakers

How fast are LLM inference engines anyway? · Taking Reinforcement Learning Cross Datacenter

Prime Intellect

5 talks · 2 speakers

Local Models: Trust, Control, Optimization · Modern Post-Training: A Deep Dive — Will Brown, Prime Intellect

Together AI

5 talks · 4 speakers

Agents at Scale: Inside MiniMax's Model and the Infrastructure Behind It · Engineering voice agents: Latency, quality, and scale

Windsurf

5 talks · 3 speakers

Agents are built at the fringe: getting from 90 to 100 · Embeddings are Stunting Agents: How Codeium Breaks Through the Ceiling for Retrieval

Amazon Web Services

4 talks · 4 speakers

Building Agents with Amazon Nova Act and MCP — Du’An Lightfoot and Banjo Obayomi · Ship it! Building Production-Ready Agents

Anterior

4 talks · 3 speakers

Don't be data poor · Make your LLM app a Domain Expert: How to Build an LLM-Native Expert System

Baseten

4 talks · 4 speakers

From model weights to API endpoint with TensorRT-LLM · Introduction to LLM serving with SGLang

Cline

4 talks · 3 speakers

Don't Build Slop (4 Levels of AI Agent Maturity) · Evals Are Broken, Use Them Anyway

Factory

4 talks · 2 speakers

Building Reliable Agentic Systems · How Forward Deployed Engineering is done at Factory

LangChain

4 talks · 3 speakers

3 ingredients for building reliable enterprise agents · Architecting and Testing Controllable Agents

Notion

4 talks · 3 speakers

How to build world-class AI products — Sarah Sachs (Notion) and Carlos Esteban (Braintrust) · Notion's Token Town

Pydantic

4 talks · 1 speaker

From Stateless Nightmares to Durable Agents · Human seeded Evals — Samuel Colvin, Pydantic

Qodo

4 talks · 2 speakers

The Last Human Code Review: Building Trust in AI-Generated Code · The State of AI Code Quality: Hype vs. Reality

Sentry

4 talks · 4 speakers

Comprehend First, Code Later: The AI Skill I Rely On Daily · MCP Is Not Good Yet — David Cramer, Sentry

Snorkel AI

4 talks · 3 speakers

From Agent Traces to Agent Simulations — Rustem Feyzkhanov, Snorkel AI · Stop Making Models Bigger, Make Them Behave — Kobie Crawford, Snorkel

Sourcegraph

4 talks · 3 speakers

Building AI Agents with Real ROI in the Enterprise SDLC · The AI emperor has no DAUs: why most devs still don't use code AI

Unsloth

4 talks · 1 speaker

Compression at the Edge · Fixing bugs in Gemma, Llama & Phi-3

AI Hero

3 talks · 1 speaker

Full Walkthrough: Workflow for AI Coding — Matt Pocock · Building Great Agent Skills: The Missing Manual

Amazon AGI Lab

3 talks · 3 speakers

Amazon AGI · From RL to IRL — Gaurav Mishra, Amazon AGI Lab

Amplify Partners

3 talks · 1 speaker

Frontier Feud · The 2025 AI Engineering Report — Barr Yaron, Amplify Partners

Applied Compute

3 talks · 4 speakers

Bringing Continual Learning into Enterprises · Efficient Reinforcement Learning

Bright Data

3 talks · 2 speakers

From MCP to Scale: Pipelines That Build Themselves · The Rise of CaaS: Context-as-a-Service for Agentic AI

Cohere

3 talks · 3 speakers

Building enterprise LLM agents that work · Building SOTA Open Weights Tool Use: The Command R Family

Databricks

3 talks · 2 speakers

From Chaos to Choreography: Multi-Agent Orchestration Patterns That Actually Work — Sandipan Bhaumik · On Engineering AI Systems that Endure The Bitter Lesson

Elastic

3 talks · 2 speakers

Agentic Search for Context Engineering · Information Retrieval from the Ground Up

Evil Martians

3 talks · 2 speakers

Don't just slap on a chatbot: building AI that works before you ask · GTM Is You - Victoria Melnikova, Evil Martians

Hex

3 talks · 2 speakers

Hiring & Building an AI Engineering Team · The AI Pivot: With Chris White of Prefect & Bryan Bischof of Hex

Hypermode

3 talks · 3 speakers

Git push, get an AI API. · Hypermode Launch

IBM

3 talks · 3 speakers

Harnesses in AI: A Deep Dive · OpenRAG: An open-source stack for RAG — Phil Nash

Keycard

3 talks · 4 speakers

How to Secure Agents using OAuth · It's 10pm. Do You Know Where Your Agents Are?

Meta

3 talks · 3 speakers

Building Deterministic Infrastructure for Non-Deterministic AI Agents · Production Evals For Agentic AI Systems

Netflix

3 talks · 3 speakers

AI Agents for Performance: Ship Faster, Pay Less · The Infinite Software Crisis

OpenPipe

3 talks · 1 speaker

How to Train Your Agent: Building Reliable Agents with RL · How we scaled 500m AI agents in production with 2 engineers

OpenRouter

3 talks · 1 speaker

fun stories from building OpenRouter and where all this is going · The Next Unicorns: 7 Top AI startups from the HF0 Residency

Orb

3 talks · 2 speakers

Monetizing AI — Alvaro Morales, Orb · Revenue Engineering: How to Price (and Reprice) Your AI Product

Parlance Labs

3 talks · 1 speaker

How To Build an AI Strategy That Fails · How to construct domain-specific LLM evaluation systems.

Poolside

3 talks · 5 speakers

AGI: The Path Forward · The Messy Reality of Scale: Synthetic Data and Pre-Training

Raindrop

3 talks · 3 speakers

Building AI Products That Actually Work · Designing Agents (The Floor Is the Frontier)

Red Hat

3 talks · 3 speakers

Lobster Trap: OpenClaw in Containers from Local to K8s and Back · Strategies for LLM Evals (GuideLLM, lm-eval-harness, OpenAI Evals Workshop) — Taylor Jordan Smith

Roboflow

3 talks · 3 speakers

How Transformers Finally Ate Vision · State of the Union: Why Local, Why Now

Sierra

3 talks · 3 speakers

Rise of the AI Architect — Clay Bavor and Alessio Fanelli · The Agent Development Life Cycle

Sonar

3 talks · 3 speakers

Can LLMs generate Enterprise Quality Code? — Prasenjit Sarkar, Sonar · Guide, Verify, Solve: The Engineering Discipline Agentic Development Demands

Supabase

3 talks · 2 speakers

Combine Skills and MCP to Close the Context Gap · Skill Issue: How We Used AI to Make Agents Actually Good at Supabase

Temporal

3 talks · 2 speakers

Building Durable, Production-Ready Agents with OpenAI SDK and Temporal · MCP Tasks (async)/ Why the heck aren't any agents supporting MCP tasks/async?

The Pragmatic Engineer

3 talks · 1 speaker

Building [REDACTED:username]: Gergely Orosz × Simon Eskildsen · Software Engineering + AI = ?

Towards AI

3 talks · 3 speakers

Build Your Own Deep Research Agent + Technical Writer · Context Engineering in 2026: Compaction, Memory & Cost

Unblocked

3 talks · 3 speakers

Building agents is trivial now, context is the next frontier · Mergeable by default: Building the context engine to save time and tokens

Vectara

3 talks · 1 speaker

Enterprise Deep Research: The Next Killer App for Enterprise AI — Ofer Mendelevitch, Vectara · open-rag-eval: RAG Evaluation without "golden" answers.

Weights & Biases

3 talks · 2 speakers

Judging LLMs · Productionizing GenAI Models – Lessons from the world's best AI teams

Zapier

3 talks · 4 speakers

How Zapier Builds AI Products and Features With the Help of Braintrust · Turning Fails into Features: Zapier’s Hard-Won Eval Lessons

Zed

3 talks · 3 speakers

Building an ACP-Compatible Agent Live — Bennet Fenner, Zed · CI in the Era of AI: From Unit Tests to Stochastic Evals

Agentic AI Foundation

2 talks · 1 speaker

Build Systems, Not Code · Building an Autonomous Engineering Org

Agentuity

2 talks · 1 speaker

Conquering Agent Chaos · The Agent-Native Company

AI21 Labs

2 talks · 2 speakers

RAG Evaluation Is Broken! Here's Why (And How to Fix It) · Real AI Agents Need Planning, Not Just Prompting

Alithea Bio

2 talks · 1 speaker

Agents Need Receipts, Not More Tool Calls · Agents Need Receipts, Not More Tool Calls

AlixPartners

2 talks · 2 speakers

DSPy: The End of Prompt Engineering · The Billable Hour is Dead; Long Live the Billable Hour?

Arcee AI

2 talks · 2 speakers

Local Models: Trust, Control, Optimization · The Base Model is Dead

Bespoke Labs

2 talks · 2 speakers

Data and Environment Curation for Post-training LLMs · OpenThoughts: Data Recipes for Reasoning Models

Bismuth

2 talks · 2 speakers

Agents reported thousands of bugs, how many were real? - Ian Butler and Nick Gregory · How to Improve your Vibe Coding — Ian Butler

Bloomberg

2 talks · 2 speakers

Challenges to Scaling Agents for Generative AI Products · What We Learned Deploying AI within Bloomberg’s Engineering Organization

Browserbase

2 talks · 1 speaker

Bringing agents onto the world wide web · The Web Browser Is All You Need

Cartesia

2 talks · 2 speakers

Serving Voice AI at Scale — Arjun Desai (Cartesia) & Rohit Talluri (AWS) · State Space Models for Realtime Multimodal Intelligence

Catio

2 talks · 3 speakers

AI Copilots for Tech Architecture: The Highest-ROI Use Case You’re Not Building — Boris Bogatin and Toufic Boubez, Catio · Grounded Reasoning Systems for Cloud Architecture

Cerebras

2 talks · 3 speakers

Fast Models Need Slow Developers · From Mixture of Experts to Mixture of Agents … with Super Fast Inference

Character.ai

2 talks · 2 speakers

Evaling Video Slop · The Hierarchy of Needs for Training Dataset Development

Chroma

2 talks · 2 speakers

How to look at your data; what to look for, how to measure · Retrieval Augmented Generation in the Wild

Cognition

2 talks · 2 speakers

Devin 2.0 and the Future of SWE · The State of Model Routing — NVIDIA, Cognition, OpenRouter

Comet ML / OpenClaw

2 talks · 1 speaker

Dark Factory: OpenClaw Ships Faster Than You Can Read the Diff · Malleable Evals: Why Are We Still Evaluating Adaptive Systems with Static Tests?

Contextual AI

2 talks · 3 speakers

Forget RAG Pipelines—Build Production-Ready AI Agents in 15 Minutes · Specialized RAG Agents: Lessons learned from deploying complex AI systems in production

Convex

2 talks · 2 speakers

Building an AI assistant that makes phone calls · Convex Launch

DAgger

2 talks · 3 speakers

Containing Agent Chaos · Ship Agents that Ship: A Hands-On Workshop for SWE Agent Builders

Datadog

2 talks · 2 speakers

The Devops Engineer Who Never Sleeps · Why Your Agent Disagrees With Itself (And What To Do About It)

Decoding AI

2 talks · 1 speaker

Build Your Own Deep Research Agent + Technical Writer · Turn 10,994 Notes Into Your Agents' Memory

Deepgram

2 talks · 2 speakers

Building & Scaling an AI Agent Swarm of low latency real time voice bots! · Giving a Voice to AI Agents

eBay

2 talks · 1 speaker

Agents Need Feature Flags · ReviewDebt: a practical framework for scoring every pull request — Sachin Gupta, eBay

Every

2 talks · 2 speakers

Benchmarks Are Memes: How What We Measure Shapes AI—and Us · Dispatch from the Future: building an AI-native Company – Dan Shipper, Every, AI & I

EXO Labs

2 talks · 1 speaker

Frontier AI at Home (literally) · State of the Union: Why Local, Why Now

FactSet

2 talks · 1 speaker

How to Build Planning Agents Without Losing Control - Yogendra Miraje, FactSet · Skills are new features: Building Skill-Centric Harness — Yogendra Miraje, FactSet

Featherless.ai

2 talks · 2 speakers

The Next Unicorns: 7 Top AI startups from the HF0 Residency · WTF do people use Open Models for??

Fireworks AI

2 talks · 1 speaker

Customized, production ready inference with open source models: Dmytro (Dima) Dzhulgakov · Making Open Models 10x faster and better for Modern Application Innovation

Freeplay

2 talks · 1 speaker

Build Dynamic Products, and Stop the AI Sideshow · The Build-Operate Divide: Bridging Product Vision and AI Operational Reality

Google Labs

2 talks · 2 speakers

Proactive Agents · Your Coding Agent Just Got Cloned And Your Brain Isn't Ready

Graphite

2 talks · 1 speaker

AI-powered entomology: Lessons from millions of AI code reviews · Don’t get one-shotted: Use AI to test, review, merge, and deploy code — Tomas Reimers, Graphite

Hasura

2 talks · 2 speakers

Building efficient hybrid context query for LLM grounding · Hasura Launch: Realtime Data Connectivity for AI

Intercom

2 talks · 2 speakers

How Building with AI Can Double the Throughput of Your Engineering Team · Shipping an Enterprise Voice AI Agent in 100 Days

Intuit

2 talks · 2 speakers

How Intuit uses LLMs to explain taxes to millions of taxpayers · Why Off-the-Shelf AI Doesn't Understand Money

Kepler

2 talks · 2 speakers

How Forward Deployed Engineering is done [REDACTED:username] Kepler · How Kepler Built Verifiable AI for Financial Services

Krea

2 talks · 1 speaker

Perceptual Evaluations: Evals for Aesthetics — Diego Rodriguez, Krea.ai · The Next Unicorns: 7 Top AI startups from the HF0 Residency

Krea.ai

2 talks · 2 speakers

Infra behind Krea 2 - How to train and serve at scale · Training Krea 2 - What matters in generative model training.

LanceDB

2 talks · 1 speaker

Scaling Enterprise-Grade RAG Systems: Lessons from the Legal Frontier · The Hierarchy of Needs for Training Dataset Development

Linear

2 talks · 2 speakers

Building the platform for agent coordination · Taste & Craft: A Conversation with Tuomas Artman, CTO of Linear, and Gergely Orosz of The Pragmatic Engineer

LinkedIn

2 talks · 3 speakers

360Brew: LLM-based Personalized Ranking and Recommendation — Hamed Firooz and Maziar Sanjabi, LinkedIn AI · Lessons from Building LinkedIn's GenAI Platform

Liquid AI

2 talks · 1 speaker

Everything I Learned Training Frontier Small Models · Everything you need to know about Finetuning and Merging LLMs

Log10

2 talks · 2 speakers

Best Practices for Evaluating Large Language Model Applications with llmeval: Niklas Nielsen · What It Actually Takes to Deploy GenAI Applications to Enterprises

Mastra

2 talks · 1 speaker

Agents vs Workflows: Why Not Both? · Every Harness Will Become A Claw

METR

2 talks · 1 speaker

Long Tasks and Experienced Open Source Dev Productivity · Why Agent Hype can fall short of reality – Joel Becker, METR

Microsoft Research

2 talks · 2 speakers

GraphRAG methods to create optimized LLM context windows for Retrieval — Jonathan Larson, Microsoft · UX Design Principles for (Semi) Autonomous Multi-Agent Systems

MiniMax

2 talks · 1 speaker

Agents at Scale: Inside MiniMax's Model and the Infrastructure Behind It · Minimax M2

Mistral AI

2 talks · 2 speakers

Decoding Mistral AI's Large Language Models · Why TTS Models Now Look Like LLMs — Samuel Humeau, Mistral

Morgan Stanley

2 talks · 2 speakers

ALPHALAB: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs — Brendan Rappazzo · What RL Means for Agents

Novartis

2 talks · 1 speaker

Agentic Enterprise: What Your CEO Must Know About AI · LLM Scientific Reasoning: How to Make AI Capable of Nobel Prize Discoveries

Nubank

2 talks · 2 speakers

Simulation-Maxxing: How Nubank ships agents 20× faster with simulations · We Vetted 2,000 AI Skills Before They Reached Developers

Oleve

2 talks · 1 speaker

Building AI Products That Actually Work · The New Lean Startup

Osmantic

2 talks · 1 speaker

State of the Union: Why Local, Why Now · The Desktop Frontier — Ahmad Osman, Osmantic

Paperclip

2 talks · 1 speaker

Paperclip: Open Source Human Control Plane for AI Labor — Dotta · What Does Done Even Mean? Agents and Paperclip's Liveness Model - Dotta, Paperclip

Pi Labs

2 talks · 1 speaker

Building Metrics That Actually Work — David Karam, Pi Labs · Layering every technique in RAG, one query at a time

Pinterest

2 talks · 3 speakers

Medic for Apache Spark - First Aid for Failing Jobs - Drasko Profirovic, Pinterest · What We Learned from Using LLMs in Pinterest

PostHog

2 talks · 2 speakers

LLM codegen fails and how to stop 'em · Self Driving Products: Product Signals to Pull Requests

Prefect

2 talks · 2 speakers

The AI Pivot: With Chris White of Prefect & Bryan Bischof of Hex · Your MCP Server is Bad and You Should Feel Bad

PromptHub

2 talks · 1 speaker

Prompt Engineering Tactics · The Model Isn’t Wrong—You’re Just Bad at Prompting

Ramp

2 talks · 2 speakers

How Forward Deployed Engineering is done at Ramp · Scaffold Wisely

Reflection AI

2 talks · 2 speakers

Frontier Feud · RL for Autonomous Coding — Aakanksha Chowdhery, Reflection AI

Replit

2 talks · 2 speakers

Building AI For All · The 3 Pillars of Autonomy – Michele Catasta, Replit

Rexmore

2 talks · 2 speakers

The Cure for the Vibe Coding Hangover · The Dark Arts of Web Automation: Teaching Agents to Use Websites Like Humans

Root Signals

2 talks · 1 speaker

Agent Evals: Finally, With The Map · Will Agent Evaluation via MCP Stabilize Agent Networks?

RunPod

2 talks · 1 speaker

GPU Cloud Deployment Without Leaving Your IDE — Audry Hsu, RunPod · Under 5 minutes to a deployed LLM endpoint — Audry Hsu, RunPod

Safe Intelligence

2 talks · 2 speakers

BDD, ADR, PRD, WTF: Capturing Decisions for Humans and AI Alike — Michal Cichra, Safe Intelligence · Spec-Driven Testing for Agents With A Brain the Size of A Planet — Steven Willmott, Safe Intelligence

SambaNova Systems

2 talks · 3 speakers

Build enterprise generative AI apps using Llama-3 at 1,000 tokens/s on the SambaNova AI platform · Llama 3 at 1[REDACTED:password]000 tok/s on the SambaNova AI Platform

SemiAnalysis

2 talks · 1 speaker

Compute & System Design for Next Generation Frontier Models · The Geopolitics of AI Infrastructure

Sizzy

2 talks · 1 speaker

From Vibe Coding to Vibe Engineering · The End of Apps

Snyk

2 talks · 3 speakers

Agentic Development Security · Through the AI Fog: The architectural decision the next 24 months of agentic security depends on.

Sourcegraph/Amp

2 talks · 2 speakers

2026: The Year the IDE Died · The emerging skillset of wielding coding agents

Stanford University

2 talks · 1 speaker

Does AI Actually Boost Developer Productivity? (Stanford / 100k Devs Study) · How to Quantify AI ROI in Software Engineering (Stanford Study / 120k Devs)

Stripe

2 talks · 2 speakers

Building safe Payment Infrastructure for the autonomous economy · Mastering AI Pricing — Mayank Pant, Stripe

Temporal Technologies

2 talks · 2 speakers

Events are the Wrong Abstraction for Your AI Agents · Vision: Zero Bugs

The Browser Company

2 talks · 2 speakers

From Arc to Dia: Lessons learned in building AI Browser · Prototyping as Leadership: How a CTO Ships with AI Agents

Thomson Reuters

2 talks · 2 speakers

From Copilot to Colleague: Building Trustworthy Productivity Agents for High-Stakes Work · Missing pieces of workflow automation

tldraw

2 talks · 1 speaker

Agents on the Canvas in tldraw · tldraw computer

Traceloop

2 talks · 1 speaker

OpenLLMetry is all you need · Prompt Engineering is Dead

Trelis Research

2 talks · 1 speaker

MCP Agent Fine-Tuning Workshop - Ronan McGovern · Text-to-Speech Data Preparation and Fine-tuning Workshop - Ronan McGovern

Twilio

2 talks · 3 speakers

Cooking with fire without burning down the kitchen · The Robots Are Coming for Your Job, and That's Okay

Uber

2 talks · 4 speakers

Agentic SDLC at Uber - Building Blocks for Uber’s Software Factory · Building Closed-Loop Evals for a Multimodal Agent at Uber Scale

UC Berkeley

2 talks · 2 speakers

Beyond Static Intelligence: Evaluating Continual Learning · What We Learned From A Year of Building With LLMs

UCAL Berkeley

2 talks · 1 speaker

Anthropic's CCA Exam as a Field-Guide for Agentic Engineering · Why Agentic Systems Need Ontologies

Upside

2 talks · 2 speakers

How Juries and Librarians Can Solve GTM's AI Trust Problem · The Next Unicorns: 7 Top AI startups from the HF0 Residency

Warp

2 talks · 1 speaker

ChatGPT is poorly designed. So I fixed it · LLM Knowledge Bases: a practical guide

Wisedocs

2 talks · 1 speaker

Benchmarking Coding Agents on New vs Legacy Codebases — Denys Linkov, Wisedocs · Structuring a modern AI team

Wordware

2 talks · 2 speakers

Beyond Conversation: Why Documents Transform Natural Language into Code · Just do it. (let your tools think for themselves) - Robert Chandler

Writer

2 talks · 2 speakers

Building Trust in Enterprise AI: Evaluating Domain-Specific LLMs for Real-World Financial Scenarios · When Vectors Break Down: Graph-Based RAG for Dense Enterprise Knowledge

Y Combinator

2 talks · 2 speakers

Every company should have a Brain — Garry Tan, Y Combinator · Imagination Engineering

Yutori

2 talks · 2 speakers

Computer-use models will agentify the web, not APIs · The Bitter Layout or: How I Learned to Love the Model Picker

Zep

2 talks · 1 speaker

Memory Masterclass: Make Your AI Agents Remember What They Do! — Mark Bain, AIUS · Stop Using RAG as Memory

.txt (Outlines)

1 talk · 1 speaker

No more bad outputs with structured generation

[REDACTED:username]

1 talk · 1 speaker

Knowledge Graphs & GraphRAG: Techniques for Building Effective GenAI Applications

11X

1 talk · 2 speakers

Building Alice’s Brain: an AI Sales Rep that Learns Like a Human - Sherwood & Satwik, 11x

14.ai

1 talk · 1 speaker

Building Reliable Support Agents Using the Effect TypeScript Library - Michael Fester

8th Light

1 talk · 1 speaker

The Coherence Trap: Why LLMs Feel Smart (But Aren’t Thinking)

Ably

1 talk · 1 speaker

Why Your AI UX Is Broken (and It's Not the Model's Fault)

Abridge

1 talk · 1 speaker

From Ambient Documentation to Clinical Intelligence

Abundant

1 talk · 1 speaker

Agents are Robots Too: What Self-Driving Taught Me About Building Agents — Jesse Hu, Abundant

Abundant AI

1 talk · 1 speaker

SWE-Marathon: Evaluating Coding Agents at Billion-Token Scale - Rishi Desai, Abundant AI

Accenture

1 talk · 2 speakers

Most Enterprise Agentic Projects Are Doomed — Here’s Why

Adaption

1 talk · 1 speaker

Adaption Labs — Gradient-Free Continual Learning

Adaptive ML

1 talk · 1 speaker

Scaling Reinforcement Learning: Lessons from Trillion-Token Deployments at Fortune 500s

Adept

1 talk · 1 speaker

Climbing the Ladder of Abstraction

Adobe

1 talk · 1 speaker

How to Run Evals at Scale: Thinking Beyond Accuracy or Similarity

Aech AI

1 talk · 1 speaker

Privacy First Enterprise AI: Building AI Agents that Never Leave Your Security Boundary

Agenta

1 talk · 1 speaker

Judge the Judge: Building LLM Evaluators That Actually Work with GEPA — Mahmoud Mabrouk, Agenta AI

Agnostiq (Covalent)

1 talk · 1 speaker

Covalent Launch: The GPU Cheatcode: Fine-tune 20 Llama Models in 5 Minutes

AI Snake Oil

1 talk · 1 speaker

Building and evaluating AI Agents That Matter

AIUS Technologies

1 talk · 1 speaker

Memory Masterclass: Make Your AI Agents Remember What They Do! — Mark Bain, AIUS

Alibaba Group / Qwen

1 talk · 1 speaker

The Future of Qwen: A Generalist Agent Model

All Hands AI

1 talk · 1 speaker

The Many Ends of Programming

All Hands AI / OpenHands

1 talk · 1 speaker

Software Development Agents: What Works and What Doesn't

Allen Institute for AI (Ai2); Interconnects.ai

1 talk · 1 speaker

A Taxonomy for Next-Generation Reasoning Models

AlleyCorp

1 talk · 1 speaker

Shipping something to someone always wins

AllHands

1 talk · 1 speaker

Automating Large-Scale Refactors with Parallel Agents

Allos AI

1 talk · 1 speaker

Trading Desks to Clinical Trials: Parallels in Applied Vertical AI

Alma

1 talk · 1 speaker

My AI Thinks I'm Eating My Feelings (and Other Nutritional Insights)

Alpic

1 talk · 1 speaker

Why MCP and ChatGPT Apps Use Double Iframes — Frédéric Barthelet, Alpic

Altos Labs

1 talk · 1 speaker

From Tokens to Cells: Foundation Models for Single-Cell Biology - Akram Baharlouei, Altos Labs

Amazon AGI SF Lab

1 talk · 1 speaker

Useful General Intelligence

Ambient

1 talk · 1 speaker

Harnessing the Power of LLMs Locally

Amp Code / Sourcegraph

1 talk · 1 speaker

Amp Code: Next-Generation AI Coding

Amplifon

1 talk · 1 speaker

One Registry to Rule them All - Sonny Merla, Mauro Luchetti, & Mattia Redaelli, Quantyca

Andon Labs

1 talk · 1 speaker

Vending-Bench: Long-Horizon Agent Evals — Lukas Petersson, Andon Labs

Annicha Labs

1 talk · 1 speaker

Beyond the Harness: A Journey Towards Adaptive Engineering - Rajiv Chandegra, Annicha Labs

Apify

1 talk · 1 speaker

The rise of the agentic economy on the shoulders of MCP

ARC Prize Foundation

1 talk · 1 speaker

Measuring AGI: Interactive Reasoning Benchmarks

Arcjet

1 talk · 1 speaker

How to defend your sites from AI bots

Arena.ai

1 talk · 1 speaker

What Do Models Still Suck At?

Ario

1 talk · 1 speaker

The Adversarial Path to the Personal Assistant

Arista Networks

1 talk · 1 speaker

How to Build Your Own AI Data Center in 2025

Arithmetic

1 talk · 1 speaker

Training Frontier Models to Out-Think Hackers — Uri Rolls, Arithmetic & Thom Wolf, Hugging Face

Arklex AI; Columbia University

1 talk · 1 speaker

How to Improve Your Agents: Academic Lit Review

Arrakis

1 talk · 1 speaker

Arrakis: How To Build An AI Sandbox From Scratch

Artificial Analysis

1 talk · 2 speakers

Trends Across the AI Frontier

Atlan

1 talk · 1 speaker

WTF Is the Context Layer? The Missing Infrastructure for Production Agents

Auditoria AI

1 talk · 1 speaker

Your Finance Agent's Bottleneck Is You

Auth0

1 talk · 2 speakers

Securing Agents with Open Standards

AutoGPT

1 talk · 3 speakers

The Future of Work

Automattic

1 talk · 1 speaker

500 people vibe-coded for 30 days. I was one of them.

Aviator

1 talk · 1 speaker

How to Kill the Code Review

AXA

1 talk · 1 speaker

Optimizing LLMs in Insurance with DSPy: Beyond Manual Tuning

Banking Circle

1 talk · 1 speaker

Platforms for Humans and Machines: Engineering for the Age of Agents — Juan Herreros Elorza

Baz

1 talk · 1 speaker

Bending a Public MCP Server Without Breaking It — Nimrod Hauser, Baz

BBD Software

1 talk · 1 speaker

Unlocking Africa's Potential with AI — Thabang Ledwaba

Bee, Amazon

1 talk · 1 speaker

Privacy-Preserving Intelligence — Steve Korshakov, Bee (acq. Amazon)

Bench Computing

1 talk · 1 speaker

A2A & MCP: Automating Business Processes with LLMs

Best Buy

1 talk · 1 speaker

Building Multi-agent Systems with Finite State Machines

Better Auth

1 talk · 1 speaker

Full Workshop: Agent Auth Protocol — Paola Estefanía de Campos, Better Auth

BetterUp

1 talk · 1 speaker

Hacking Subagents Into Codex CLI — Brian John, BetterUp

Bitly

1 talk · 1 speaker

Let’s Talk About FOMAT – Fear of Missing Agent Time

Black Forest Labs

1 talk · 1 speaker

Black Forest Labs: FLUX, Open Research, and the Future of Visual AI

BlackRock

1 talk · 2 speakers

How BlackRock Builds Custom Knowledge Apps at Scale

Block

1 talk · 1 speaker

Your AI Agent Isn't an Engineer: The Art of Thoughtful Anthropomorphism

Bolt.new / StackBlitz

1 talk · 1 speaker

Bolt.new: How we scaled $0-20m ARR in 60 days, with 15 people

booking.com

1 talk · 1 speaker

Building AI Agents with Real ROI in the Enterprise SDLC

BotDojo

1 talk · 1 speaker

BotDojo Launch: Enhancing AI Assistants with Evaluations and Synthetic Data

Boundary

1 talk · 1 speaker

fighting slop with slop

Box

1 talk · 1 speaker

Building an Agentic Platform

Brightwave

1 talk · 1 speaker

Trust, but Verify: High-Fidelity Reasoning in Agentic Workflows

Bugcrowd; Carnegie Mellon University

1 talk · 1 speaker

Teaching AI to Find Real Vulnerabilities — Prof. David Brumley, Bugcrowd

Callosum

1 talk · 1 speaker

Scaling the Next Paradigm of Heterogeneous Intelligence

Callstack

1 talk · 1 speaker

OpenClaw in Your Hand: Building a Physical AI Terminal for Local LLM Agents

Capital One

1 talk · 1 speaker

Developer Experience in the Age of AI Coding Agents

Casco

1 talk · 1 speaker

How we hacked YC Spring 2025 batch’s AI agents

Caylent

1 talk · 1 speaker

POC to PROD: Hard Lessons from 200+ Enterprise GenAI Deployments

Checkout.com

1 talk · 1 speaker

Your coding agent doesn't always follow your rules

Cherrypick

1 talk · 1 speaker

Ralph Loops: Build Dumb AI Loops That Ship

Chime

1 talk · 1 speaker

The Build-Operate Divide: Bridging Product Vision and AI Operational Reality

China Resources Holdings

1 talk · 1 speaker

Build for the Memo, Not the Demo — Notes from 200 Investment Committees

Circle

1 talk · 1 speaker

Automating Escrow with USDC and AI

Cisco / Outshift by Cisco

1 talk · 1 speaker

Multi-Agent AI and Network Knowledge Graphs for Change Management and Network Testing

CloudChef

1 talk · 1 speaker

General purpose robots as professional Chefs

cmpnd

1 talk · 1 speaker

The Unreasonable Effectiveness of Separating the Task from the Model

CodiumAI

1 talk · 1 speaker

Move Fast Break Nothing

Cognee

1 talk · 1 speaker

Memory Masterclass: Make Your AI Agents Remember What They Do! — Mark Bain, AIUS

Cognition (Devin)

1 talk · 1 speaker

The Making of Devin

Cognition AI

1 talk · 1 speaker

How Forward Deployed Engineering is done at Cognition

Comfy Org

1 talk · 2 speakers

ComfyUI Workshop with ComfyAnonymous and Jedrick Kosinski

Conductor

1 talk · 1 speaker

Content Is Code

Confident Security

1 talk · 1 speaker

The Unofficial Guide to Apple’s Private Cloud Compute

Conviction

1 talk · 1 speaker

State of Startups and AI 2025

Corridor

1 talk · 1 speaker

The AI bugpocalypse is here. Now what?

CoupleWork AI

1 talk · 2 speakers

AI is the World’s largest Relationship Therapist — Clay Cockrell & Tony Fabrikant, CoupleWork AI

Coval

1 talk · 1 speaker

From Self-driving to Autonomous Voice Agents — Brooke Hopkins, Coval

crewAI

1 talk · 1 speaker

Using agents to build an agent company

Crusoe

1 talk · 1 speaker

Accelerating Mixture of Experts Training With Rail-Optimized InfiniBand Networking in Crusoe Cloud

CRV

1 talk · 1 speaker

The AI Pivot: With Chris White of Prefect & Bryan Bischof of Hex

Cua

1 talk · 3 speakers

Computer-Use 2.0: Agents Just Got Multi-Cursor

DataChain

1 talk · 1 speaker

When Agents Meet Physical Data: The Other Physics of Agent Harnesses

Datacurve

1 talk · 1 speaker

DeepSWE: A Contamination-Resistant Coding Benchmark — James Shi, Datacurve

Datalab

1 talk · 1 speaker

Small AI Teams with Huge Impact — Vik Paruchuri, Datalab

DataRobot

1 talk · 1 speaker

Skills are the New SDKs

Datasette

1 talk · 1 speaker

Claude Fable, Claude Tag, and Anthropic's Culture — Cat Wu & Thariq Shihipar ft Simon Willison

DatologyAI

1 talk · 1 speaker

Data Quality is the Compute Multiplier

Dawn Analytics

1 talk · 1 speaker

The era of unbounded products: Designing for Multimodal I/O

Daytona

1 talk · 1 speaker

AX is the only Experience that Matters

dbt Labs

1 talk · 1 speaker

AI’s Jurassic Park Period

Decagon

1 talk · 1 speaker

How Forward Deployed Engineering is done at Decagon

Decawork

1 talk · 1 speaker

IT Admin for the AI Workforce — Sarthak Aggarwal, Decawork

DeepMind

1 talk · 1 speaker

Unveiling the latest Gemma model advancements

deepset

1 talk · 1 speaker

Let LLMs Wander: Engineering RL Environments — Stefano Fiorucci

deepset GmbH

1 talk · 1 speaker

What Breaks When You Build AI Under Sovereignty Constraints

Deno

1 talk · 1 speaker

Security Firewall for Agents

DevDay

1 talk · 1 speaker

How to Hire AI Engineers When Everyone Is Cheating With AI

Discord

1 talk · 1 speaker

Iterating on LLM apps at scale: Learnings from Discord

Docker

1 talk · 1 speaker

Unlock Agent Autonomy: The Runtime for AI-Native Systems

DSPy

1 talk · 1 speaker

The Unreasonable Effectiveness of Separating the Task from the Model

Duolingo

1 talk · 1 speaker

Build AI Systems for Discernment, Not Approval - Angel Ortmann Lee, Duolingo

DX

1 talk · 1 speaker

Leadership in AI-Assisted Engineering

Dylibso

1 talk · 1 speaker

The State of MCP Observability: Observable.tools — Alex Volkov and Benjamin Eckel, Weights & Biases and Dylibso

E2B

1 talk · 1 speaker

How to add secure code interpreting in your AI app

Earendil

1 talk · 2 speakers

The Friction Is Your Judgment

Echo AI

1 talk · 1 speaker

What It Actually Takes to Deploy GenAI Applications to Enterprises

Effectful Technologies Inc

1 talk · 1 speaker

Vibe Engineering Effect Apps

Emergence

1 talk · 1 speaker

Emergence Launch: AI Agents and the future enterprise

Emulated

1 talk · 2 speakers

Emulated: The data for fully autonomous software engineers and companies

Engram

1 talk · 1 speaker

Scaling Compute on Context

Ensemble Health Partners

1 talk · 1 speaker

AI That Pays: Lessons from Revenue Cycle

Entry Point AI

1 talk · 1 speaker

No-code Fine-tuning: Mark Hennings

EpicAI.pro

1 talk · 1 speaker

Letting AI Interface with Your App with MCP

Erel Labs

1 talk · 1 speaker

MCP-UI: Extending the Frontier — Liad Yosef and Ido Salomon, MCP Apps

Etsy

1 talk · 1 speaker

What if the harness mattered more than the model? - Aditya Bhargava, Etsy

Every/Cora

1 talk · 1 speaker

The Era of Compound Engineering

exa

1 talk · 1 speaker

Building a Smarter AI Agent with Neural RAG

EyeLevel.ai

1 talk · 1 speaker

EyeLevel Launch: Your RAG is Tripping, Here's the Real Reason Why

Factory AI

1 talk · 1 speaker

Making Codebases "Agent-Ready"

FAIR, Meta

1 talk · 1 speaker

Code World Model: Building World Models for Computation

Fal

1 talk · 1 speaker

The State of Generative Media Today

Fidelity Investments

1 talk · 1 speaker

Wearing the Agent: Engineering a Family-and-Friends Personal Agent, from Group Chats to Glasses

Filed

1 talk · 1 speaker

Chat and citations won't save your vertical AI

Fireworks

1 talk · 1 speaker

Making Open Models 10x faster and better for Modern Application Innovation

Fixie.ai

1 talk · 1 speaker

Building Reactive AI Apps

Flatfile

1 talk · 1 speaker

Form factors for your new AI coworkers

Flinn AI

1 talk · 1 speaker

What the Best Agents Share

FlyersSoft

1 talk · 1 speaker

Let's integrate AI Agents in Event-Sourced Systems

Forestwalk Labs

1 talk · 1 speaker

Voice In, Visuals Out: The Agony and the Ecstasy

Form3

1 talk · 1 speaker

We Gave an Agent Production Code Access and Then Tried to Sleep at Night

Forward Future

1 talk · 1 speaker

State of the Union: Why Local, Why Now

Fractional AI

1 talk · 1 speaker

Voice Agents: the good, the bad, and the ugly

Freeman & Forrest

1 talk · 1 speaker

To the moon! Navigating deep context in legacy code with Augment Agent

Fujitsu North America

1 talk · 1 speaker

VoiceOps-fying Low-Latency Intelligence Extraction from Messy Audio Streams — Dippu Kumar Singh

Funstage GmbH

1 talk · 1 speaker

Backlog.md: Terminal Kanban Board for Managing Tasks with AI Agents — Alex Gavrilescu, Funstage

G2i

1 talk · 1 speaker

Benchmarks: The Good, the Bad, and the Ugly

Gabber

1 talk · 2 speakers

Serving Voice AI at $1/hr: Open-source, LoRAs, Latency, Load Balancing

Galileo

1 talk · 1 speaker

Taming Rogue AI Agents with Observability-Driven Evaluation

Gamma

1 talk · 1 speaker

Rethinking Team Building: How a 30-Person Startup Serves 50 Million Users — Grant Lee, Gamma

Gas Town

1 talk · 1 speaker

Agentic Security: Permissions, Provenance, and the Agent Supply Chain

Gates Foundation

1 talk · 1 speaker

Your Moat Is Your Data Model

GenAI Israel

1 talk · 1 speaker

The LLM Triangle: Engineering Principles for Robust AI Applications

General Reasoning

1 talk · 2 speakers

Scaling to Long Horizons

GenSX

1 talk · 1 speaker

How agents broke app-level infrastructure

Gimlet Labs

1 talk · 1 speaker

AI Kernel Generation: What's Working, What's Not, What's Next

Gitpod

1 talk · 1 speaker

Building CISO-approved agent fleet architecture

Glean

1 talk · 1 speaker

How to build Enterprise-aware agents

Glow

1 talk · 1 speaker

The Next Unicorns: 7 Top AI startups from the HF0 Residency

Good Collective

1 talk · 1 speaker

A Practitioner's Guide to Graphs - Tim Ainge, Good Collective

Goodfire

1 talk · 1 speaker

Why you should care about AI interpretability

Google / YouTube

1 talk · 1 speaker

Teaching Gemini to Speak YouTube: Adapting LLMs for Video Recommendations to 2B+ DAU

Google / YouTube Ads

1 talk · 2 speakers

How Evals and Prompts Shape Agent Behavior — Preetika Bhateja & Daniel Bump, YouTube Ads

Google Photos

1 talk · 1 speaker

Magic Editor Under the Hood: Weaving Generative AI into a Billion-User App

Gradient

1 talk · 1 speaker

Training Albatross: An Expert Finance LLM

Gradium AI

1 talk · 1 speaker

Neil Zeghidour - Voice AI: when is the "Her" moment?

Granola

1 talk · 1 speaker

Feedback Loops are All You Need

Grit

1 talk · 1 speaker

Code Generation and Maintenance at Scale

Groq

1 talk · 1 speaker

Breaking AI’s 1 Gigahertz Barrier

Growth Cyber

1 talk · 1 speaker

How to Build Trustworthy AI

Gru.ai

1 talk · 1 speaker

How Coding Agents Change Software Development Forever - Hailong Zhang

Guardrails AI

1 talk · 1 speaker

Trust, but Verify

Gumloop

1 talk · 1 speaker

Building a 10-Person Unicorn

Haize Labs

1 talk · 1 speaker

Fuzzing in the GenAI Era

Halluminate

1 talk · 2 speakers

The Current State of Browser Agents

Harvey

1 talk · 1 speaker

Scaling Enterprise-Grade RAG Systems: Lessons from the Legal Frontier

Heroku

1 talk · 2 speakers

Building Agentic Applications with Heroku Managed Inference and Agents — Julián Duque and Anush DSouza

Hey AI

1 talk · 1 speaker

Your AI Product Will Fail Unless You Can Explain It

HeyGen

1 talk · 1 speaker

HTML Is All Agents Need

HiddenLayer

1 talk · 1 speaker

Building security around ML

Higharc

1 talk · 1 speaker

Research to Reality: Bringing frontier ML research to production

Hinge Health

1 talk · 1 speaker

Guardrails First: Engineering Member-Facing Health AI

Hippocratic AI

1 talk · 1 speaker

200 Million Patient Interactions Later: What the Generic Voice Stack Misses

HoneyHive

1 talk · 1 speaker

Your Evals Are Meaningless (And Here’s How to Fix Them)

Hud

1 talk · 1 speaker

From Blind Spots to Merged PRs: Continuous Agentic Performance Optimization

Huge

1 talk · 1 speaker

Invisible Users, Invisible Interfaces: Accelerating Design Iteration with AI Simulation

Humanloop

1 talk · 1 speaker

Real ROI: Lessons from Enterprises that Have already succeeded with LLMs [REDACTED:username] Scale

Huxe

1 talk · 1 speaker

Everything is ugly, so go build something that isn't

Hyperbolic

1 talk · 1 speaker

Why We Don’t Need More Data Centers

Hyperspace

1 talk · 1 speaker

Hyperspace: More Nodes Is All You Need

IKEA

1 talk · 1 speaker

Build Your First Demand-Driven Context Base: Let AI Agents Tell You What They Need

Imbue

1 talk · 1 speaker

Beyond the Prototype: Using AI to Write High-Quality Code

incident.io

1 talk · 1 speaker

Lawrence Jones - Fighting AI with AI

Incubator for Artificial Intelligence (i.AI)

1 talk · 1 speaker

Why your product needs an AI product manager, and why it should be you

Independent / State of Data

1 talk · 1 speaker

State of Data

Independent Researcher

1 talk · 1 speaker

AI-Driven Multi-Document Correlation for Enterprise Financial Compliance and Fraud Detection

Inngest

1 talk · 1 speaker

Your agent architecture has a half-life of 6 months

Insight Sciences

1 talk · 1 speaker

Persona Engineering: A Field Guide to AI Synthetic Personas — Ishan Anand, InsightSciences.ai

Instacart

1 talk · 1 speaker

How Instacart transformed its search and discovery using an LLM-driven approach

Ionic Commerce

1 talk · 1 speaker

Ionic Launch: Opening the economy to AI agents

Isadora & Co | The Bloom House AI

1 talk · 1 speaker

Stop Writing Tone Instructions. Layer Them.

IT Revolution

1 talk · 1 speaker

2026: The Year the IDE Died

Iterate

1 talk · 2 speakers

Make your own event-sourced agent harness using stream processors

Jam

1 talk · 1 speaker

The AI Engineer’s Guide to Raising VC — Dani Grant (Jam), Chelcie Taylor (Notable Capital)

Jane Street

1 talk · 1 speaker

Building AI-Powered Developer Tools at Jane Street

Jedi

1 talk · 1 speaker

Reverse Conway's law and GenAI: How agents will take over the organisation

Jellyfish

1 talk · 1 speaker

What Data from 20 Million Pull Requests Reveal About AI Transformation

JoinIn AI

1 talk · 1 speaker

The Prompt Is Still a Punch Card

Jointly

1 talk · 1 speaker

The Unbearable Lightness of Agent Optimization

JP Morgan Chase

1 talk · 1 speaker

Learned Execution Graphs for Anomaly Detection & Drift in APIs — Ritvik Pandya, JP Morgan Chase

K-Scale Labs

1 talk · 1 speaker

Your Personal Open-Source Humanoid Robot for $8,999 — Jingxiang "JX" Mo, K-Scale Labs

Kalmantic Labs

1 talk · 1 speaker

Stop Renting Your Cognitive Infrastructure

Khan Academy

1 talk · 1 speaker

Scaling AI in Education: A Khanmigo case study

Kilo Code

1 talk · 1 speaker

Agentic Engineering: Working With AI, Not Just Using It — Brendan O'Leary

Klarity

1 talk · 1 speaker

E-Values: Evaluating the Values of AI

KRAFTON

1 talk · 1 speaker

I Run a Fleet of AI Agents Across Three Machines. Here's What Broke.

Langbase

1 talk · 1 speaker

Why the Best AI Agents Are Built Without Frameworks (Primitives over Frameworks)

Langbase; Command Code

1 talk · 1 speaker

Developing Taste in Coding Agents: Applied Meta Neuro-Symbolic RL — Ahmad Awais, Command Code

Langfuse

1 talk · 1 speaker

Stop Burning Tokens: Why self-improvement needs domain expertise first - Annabell Schäfer, Langfuse

Langfuse, part of ClickHouse

1 talk · 1 speaker

Skill issue: Lessons from skilling up coding agents to use Langfuse

LastMile AI

1 talk · 1 speaker

Exposing Agents as MCP Servers with mcp-agent: Sarmad Qadri

LatchBio

1 talk · 1 speaker

Verifiable Environments for AI in Biology — Kenny Workman, LatchBio

Latent Space University

1 talk · 1 speaker

AI Engineering 101

Laude Institute

1 talk · 2 speakers

Everything Is a Rollout — Alex Shaw + Ryan Marten, Terminal-Bench, Harbor, Laude Institute

Lease End

1 talk · 1 speaker

Your Fine-Tuned Model Is Tech Debt: A 50x ROI House of Cards

Legora

1 talk · 1 speaker

Agents need more than a chat

Leibniz Labs

1 talk · 1 speaker

"I've never seen anything scarier than an LLM with tool calls." — Erik Meijer aka @HeadinTheBox

LemonSlice

1 talk · 1 speaker

Voice agents with Realtime Video — Sidney Primas, LemonSlice

Lenses.io

1 talk · 2 speakers

Your Insecure MCP Server Won't Survive Production — Tun Shwe, Lenses

Lexica

1 talk · 1 speaker

On Curiosity — Sharif Shameem, Lexica

LexisNexis

1 talk · 1 speaker

Your LLM Deception Monitor Is Broken. The Fix Is in the Training Data - Sachin Kumar, LexisNexis

Lindy

1 talk · 1 speaker

The Age of the Agent

Livekit

1 talk · 1 speaker

Why ChatGPT Keeps Interrupting You

Locally AI

1 talk · 1 speaker

Running Gemma 4 On-Device: 40 Tokens/s on iPhone with MLX

Los Alamos National Laboratory

1 talk · 1 speaker

Government Agents: AI Agents Meet Tough Regulations — Mark Myshatyn, Los Alamos National Laboratory

Lovable

1 talk · 1 speaker

How Lovable self-improves every hour

Luma AI

1 talk · 1 speaker

Dream Machine: Scaling to 1m users in 4 days — Keegan McCallum, Luma AI

Luminal

1 talk · 1 speaker

Luminal - Search-Based Deep Learning Compilers - Joe Fioti

Lux Capital

1 talk · 1 speaker

Beyond the Consensus: Navigating AI’s Frontier in 2025

Lyft

1 talk · 2 speakers

Build Evals That Actually Matter - Nick Ung & Akshay Sharma, Lyft

M87 Labs

1 talk · 1 speaker

Moondream: how does a tiny vision model slap so hard?

Machinecraft

1 talk · 1 speaker

The Factory That Dreams: 39 AI Agents, No Framework

Manufact, Inc

1 talk · 1 speaker

MCP Apps: Primitives, Discovery, and the Future of Software

Manufactured

1 talk · 1 speaker

Let's Build an Agent from Scratch — Kam Lasater

Manus

1 talk · 1 speaker

Building Intelligent Research Agents with Manus

Mastercard

1 talk · 1 speaker

Navigating Challenges and Technical Debt in LLMs Deployment

Maven Clinic

1 talk · 1 speaker

How to build an AI-Native Health Company

McKinsey & Company

1 talk · 2 speakers

Moving away from Agile: What's Next?

MCP Apps

1 talk · 2 speakers

MCP Apps: Extending the Frontier — Ido Salomon & Liad Yosef

Meta PyTorch

1 talk · 1 speaker

What does it take to build a personal, local, private AI Agent that augments you deeply?

Method Financial

1 talk · 1 speaker

How we scaled 500m AI agents in production with 2 engineers

Midjourney

1 talk · 1 speaker

Second Order Effects

Mindmakers

1 talk · 1 speaker

Stop Evaluating Models Like It's the 50s - Alejandro Vidal, Mindmakers

Mistral

1 talk · 1 speaker

Decoding Mistral AI's Large Language Models

MIT Media Lab

1 talk · 1 speaker

The Agentic Web and the Bazaar Era of AI - Ramesh Raskar, MIT Media Lab

Mixedbread

1 talk · 2 speakers

How we taught agents to use good retrieval - Hanna Lichtenberg, Mixedbread AI

Modular

1 talk · 1 speaker

Unlocking Developer Productivity across CPU and GPU with MAX

Monday

1 talk · 1 speaker

From Systems of Record to Systems of Context

monday.com

1 talk · 1 speaker

From Systems of Record to Systems of Context

MongoDB / Voyage AI

1 talk · 1 speaker

RAG in 2025: State of the Art and the Road Forward

Morph Labs

1 talk · 1 speaker

The infrastructure for the singularity

Mozilla

1 talk · 2 speakers

Llamafile: bringing AI to the masses with fast CPU inference

Mozilla.ai

1 talk · 1 speaker

2025 is the Year of Evals! Just like 2024, and 2023, and …

Multinear

1 talk · 1 speaker

Practical tactics to build reliable AI apps — Dmitry Kuchin, Multinear

Muna

1 talk · 1 speaker

Compilers in the Age of LLMs

Mutagent

1 talk · 2 speakers

The Agentic AI Engineer

n8n

1 talk · 1 speaker

Building Your Own Secure AI Workflows: Human-in-the-Loop Automation with n8n

Namespace

1 talk · 1 speaker

CI/CD Is Dead, Agents Need Continuous Compute and Computers — Hugo Santos and Madison Faulkner

Nearform

1 talk · 1 speaker

Agents Building Agents

Nebius

1 talk · 1 speaker

SWE-rebench: Lessons from Evaluating Coding Agents on Real Software Engineering Tasks — Ibragim Badertdinov, Nebius

NeoCognition

1 talk · 1 speaker

Intelligence + Continual Learning = Expertise

Nereu

1 talk · 1 speaker

The Next Game Engine Won't Have a Manual

New Computer

1 talk · 2 speakers

The Intelligent Interface

New Enterprise Associates (NEA)

1 talk · 1 speaker

CI/CD Is Dead, Agents Need Continuous Compute and Computers — Hugo Santos and Madison Faulkner

New Generation (New Gen)

1 talk · 1 speaker

Machines of Buying & Selling Grace

News Corp

1 talk · 1 speaker

Stop Guessing: Build Robust AI with Layered CoT

Nori

1 talk · 1 speaker

HTML is All You Need (for Agents to Make Graphics)

Normal Computing

1 talk · 1 speaker

Going beyond RAG: Extended Mind Transformers

Northwestern Mutual

1 talk · 1 speaker

Small Bets, Big Impact: Building GenBI at a Fortune 100

Notable Capital

1 talk · 1 speaker

The AI Engineer’s Guide to Raising VC — Dani Grant (Jam), Chelcie Taylor (Notable Capital)

Notius Labs

1 talk · 1 speaker

How to Leverage Domain Expertise — Chris Lovejoy, Notius Labs

Nx

1 talk · 1 speaker

A Genius With Amnesia

OctoAI

1 talk · 2 speakers

LLM Quality Optimization Bootcamp

Ogilvy

1 talk · 1 speaker

Bypassing the Multimodal Tax: Framework-Free Hybrid RAG, Raw SQL RRF, and Live UI Telemetry

OLIVER

1 talk · 1 speaker

Bounded Autonomy: Between Free Will and Determinism

Ollama

1 talk · 1 speaker

Compression at the Edge

Omnara

1 talk · 1 speaker

The Log Is The Agent

Ona

1 talk · 1 speaker

The Missing Primitive for Agent Swarms

Onlay

1 talk · 1 speaker

Healthcare’s Agent Bytecode: X12 as the Harness for AI Agents

OpenAudio / Fish Audio

1 talk · 1 speaker

The Next Unicorns: 7 Top AI startups from the HF0 Residency

OpenClaw

1 talk · 1 speaker

I Gave an AI Agent the Keys to My Life (Here's What Happened)

OpenClaw; TextCortex at the time of recording

1 talk · 1 speaker

Scaling Agents on Kubernetes with acpx and ACP

OpenCode

1 talk · 1 speaker

AI changes *Nothing* — Dax Raad, OpenCode

OpenGov

1 talk · 1 speaker

Agents in Production: How OpenGov Built and Scaled OG Assist

OpenProse

1 talk · 1 speaker

Recursive Coding Agents

Orbis Operations

1 talk · 1 speaker

When all context matters: Extended Cache Augmented Generation (ECAG)

Orbital

1 talk · 1 speaker

Buy Now, Maybe Pay Later: Dealing with Prompt-Tax While Staying at the Frontier - Andrew Thompson

Oxylabs

1 talk · 1 speaker

How Web Data Infrastructure Powers the Next Generation of AI

Palo Alto Networks

1 talk · 1 speaker

Self-Evolving Code with AI: Enhancing Quality and Security in CI

Patho.ai

1 talk · 1 speaker

Wisdom-Driven Knowledge Augmented Generation at Scale

Perpetual

1 talk · 1 speaker

Personality-Driven Development: Exploring the Frontier of Agents with Attitude

PFF

1 talk · 1 speaker

Agents Don't Do Standups: Building the Post-Engineer Engineering Org

Pfizer

1 talk · 1 speaker

Anchoring Enterprise GenAI with Knowledge Graphs

Phaidra

1 talk · 2 speakers

Semantic Blindness: 500,000 Sensors Confused an LLM - Raahul Singh & Vanč Levstik, Phaidra

Philo Ventures

1 talk · 1 speaker

While my guitar gently speaks

Physical Intelligence

1 talk · 2 speakers

Robotics: why now?

PI

1 talk · 1 speaker

Building pi in a World of Slop

Pieces

1 talk · 1 speaker

Foundry Local: Cutting-Edge AI Experiences on Device with ONNX Runtime and Olive — Emma Ning, Microsoft

Ping Labs

1 talk · 1 speaker

Everything we knew about software has changed — Theo Browne

Polar Signals

1 talk · 1 speaker

Maximize GPU Efficiency with Continuous Profiling for GPUs

Pomerium

1 talk · 1 speaker

Claws Out: Securing and Building with OpenClaw

Portia AI

1 talk · 1 speaker

From PM at Stripe to Building an AI Startup, a Recent Founder's Journey - Mounir Mouawad

Position²

1 talk · 1 speaker

Build the AI GTM Agent That Knows the Buyer Before the First Message

Postman

1 talk · 1 speaker

Beyond Components: Designing Generative UI for MCP Apps

Prediction Guard

1 talk · 1 speaker

LLM Safeguards: Security, Privacy, Compliance, Anti-Hallucination

Privacera

1 talk · 1 speaker

Balancing Innovation with Security & Safety

Programma Labs

1 talk · 1 speaker

Computer Use at the Edge of the Statistical Precipice

Progress Software

1 talk · 1 speaker

The UX of AI: Making AI-Powered Apps Your Users Don't Hate

Project NANDA

1 talk · 1 speaker

The Agentic Web and the Bazaar Era of AI - Ramesh Raskar, MIT Media Lab

PromptLayer

1 talk · 1 speaker

How Claude Code Works

PromptQL

1 talk · 1 speaker

"Data readiness" is a Myth: Reliable AI with an Agentic Semantic Layer — Anushrut Gupta, PromptQL

Prosodica

1 talk · 2 speakers

The 100-Tool Agent Is a Trap: Scaling with Semantic Routers and JIT Context

Pruna AI

1 talk · 1 speaker

20 days of compute vs 7 hours: rethinking what state-of-the-art means — Bertrand Charpentier, Pruna AI

Pulley

1 talk · 1 speaker

Analyzing 10,000 Sales Calls with AI in 2 Weeks

pyannoteAI

1 talk · 1 speaker

Beyond Transcription: Building Voice AI That Actually Understands Conversations

Qdrant

1 talk · 1 speaker

Navigating RAG Optimization with an Evaluation-Driven Compass

Quantyca

1 talk · 2 speakers

One Registry to Rule them All - Sonny Merla, Mauro Luchetti, & Mattia Redaelli, Quantyca

Quotient

1 talk · 1 speaker

Navigating RAG Optimization with an Evaluation-Driven Compass

Quotient AI

1 talk · 2 speakers

Evaluating AI Search: A Practical Framework for Augmented AI Systems

RADiCAIT

1 talk · 1 speaker

Autonomous Agents for Scientific Tasks - Sina Shahandeh, RADiCAIT

Railway

1 talk · 1 speaker

Infra that fixes itself, thanks to coding agents — Mahmoud Abdelwahab, Railway

Rasgo

1 talk · 1 speaker

How to Build AI Agents that Actually Work

Ratel

1 talk · 1 speaker

A Song of Types and Agents

Re-skill

1 talk · 1 speaker

This video was edited with AI agent. But how?

re:ma AI

1 talk · 1 speaker

Shift Left: How to Become an AI Engineer from a Full-Stack Background

Reactor

1 talk · 1 speaker

The Next Medium: Why Real-Time Interactive Video Changes Everything — Ahmed Ahres, Reactor

Rechat

1 talk · 1 speaker

How to construct domain-specific LLM evaluation systems.

Reelful

1 talk · 1 speaker

Building an Agentic Video Editor for Mass Consumer

Ref.

1 talk · 1 speaker

Velocity Sickness: What Happens When Your Whole Team Gets 10x Faster

Reforge

1 talk · 1 speaker

Survive the AI Knife-Fight: Building Products That Win

RELAI; University of Maryland, College Park

1 talk · 1 speaker

Continual Learning for AI Agents: From Failures to Durable Improvements - Soheil Feizi, RELAI

Replicate

1 talk · 1 speaker

Design like Karpathy is watching 😎

Resolve AI

1 talk · 1 speaker

Always-on agents run production without the on-call tax

Resonate HQ, Inc

1 talk · 2 speakers

The Prompt is the Platform

Results Generation

1 talk · 1 speaker

The Miranda Hypothesis: How Hamilton (the Musical) Poisoned Your Persona Evals

Retool

1 talk · 1 speaker

How agents will unlock the $500B promise of AI

RISA Labs

1 talk · 1 speaker

Can Oncology Workflows Run Without Human Touch? - Anant Shankhdhar, Risa Labs

rtrvr.ai

1 talk · 2 speakers

Beyond APIs: How AI Web Agents Are Automating the "Long Tail" of Knowledge Work

Sakana AI

1 talk · 1 speaker

Memory Harnesses for Long-Running Research Agents

Salesforce

1 talk · 1 speaker

A Practical Guide to Efficient AI

SambaNova

1 talk · 2 speakers

Llama 3 at 1[REDACTED:password]000 tok/s on the SambaNova AI Platform

Scalekit

1 talk · 1 speaker

You Didn't Ship a Bug. You Just Wrote It for a Human.

Scorecard

1 talk · 1 speaker

The Benchmarks Game: Why It's Rigged and How You Can (Really) Win

Sematic

1 talk · 1 speaker

How to evaluate a model for your use case

SF Compute

1 talk · 1 speaker

Good design hasn’t changed with AI

SignalFire

1 talk · 1 speaker

Insights on Building AI teams

Sky Valley Ambient Computing

1 talk · 1 speaker

The Pipeline Is Dead

Smithery

1 talk · 1 speaker

Are MCPs Overhyped? A Rant about MCPs

Snapchat

1 talk · 1 speaker

Develop at Idea Velocity

SnapLogic; University of San Francisco

1 talk · 1 speaker

Breaking the Chain: Agent Continuations for Resumable AI Workflows

Snorkel

1 talk · 1 speaker

Insights from Snorkel AI running Azure AI Infrastructure

Snowglobe

1 talk · 1 speaker

Simulation-Maxxing: How Nubank ships agents 20× faster with simulations

Software 3.0, LLC

1 talk · 1 speaker

Announcing the AI Engineer Network

SonderMind

1 talk · 3 speakers

Evals Driven-Development: Engineering a Mental Health AI Coach Ethically & Safely

SpecStory

1 talk · 1 speaker

How To Build an AI Strategy That Fails

Spotify

1 talk · 1 speaker

Personalization in the Era of LLMs

Sprout Social

1 talk · 2 speakers

From Hype to Habit: How We’re Building an AI-First SaaS Company—While Still Shipping the Roadmap

StandardAgents

1 talk · 1 speaker

The Future Is Domain-Specific Agents

StarlightSearch Inc

1 talk · 1 speaker

User Signal Die at the Retrieval Boundary

Stride

1 talk · 1 speaker

Case Study + Deep Dive: Telemedicine Support Agents with LangGraph/MCP

Substrate

1 talk · 1 speaker

Substrate Launch: the API for modular AI

Super Protocol

1 talk · 1 speaker

GPU-less, Trust-less, Limit-less: Reimagining the Confidential AI Cloud

Superagentic AI

1 talk · 1 speaker

RLM: Recursive Language Models for Large Codebases

Superconductor

1 talk · 1 speaker

Multiplayer agentic engineering: enabling your whole team and your best agents to work together

SuperDial

1 talk · 1 speaker

Voice AI: Your Bot Isn't Special

Superintelligent

1 talk · 1 speaker

AI Consulting in Practice — Nathaniel Whittemore (NLW), Superintelligent

Superlinked

1 talk · 1 speaker

The Small Model Infrastructure Nobody Built (So We Did) — Filip Makraduli, Superlinked

Surge AI

1 talk · 1 speaker

When Will The Benchmaxxing Plague End?

Synth

1 talk · 1 speaker

Stateful environments for vertical agents — Josh Purtell, Synth Labs

Tailscale

1 talk · 1 speaker

What if the network was the sandbox?

Take Take Take

1 talk · 2 speakers

Building a Chess Coach

Taste Labs

1 talk · 1 speaker

Ending AI Slop

Tavily

1 talk · 1 speaker

Evaluating AI Search: A Practical Framework for Augmented AI Systems

TAVON.ai

1 talk · 1 speaker

A Piece of PI – Embedding The OpenClaw Coding Agent In Your Product

Tavus

1 talk · 1 speaker

Realtime Conversational Video with Pipecat and Tavus — Chad Bailey and Brian Johnson, Daily & Tavus

Teammates

1 talk · 1 speaker

Shipping Products When You Don’t Know What they Can Do

Telemetrak

1 talk · 2 speakers

Critical AI Inference Your CIO Can Trust

Tenex

1 talk · 1 speaker

Paying Engineers like Salespeople

Tesco

1 talk · 1 speaker

We Cut 94% of Our AI Coding Tokens With a Local Code Index. Here's the Architecture.

Tesla

1 talk · 1 speaker

Enterprise Agents Have a Structure Problem - Ishita Daga, Tesla

Tesla Optimus

1 talk · 1 speaker

Challenges in High Performance Robotics Systems

Tessl

1 talk · 1 speaker

Context Is the New Code

The New York Times

1 talk · 2 speakers

Local Agentic Theory For Mobile Games — Shafik Quoraishee & Joanne Song, The New York Times

The New York Times Games

1 talk · 1 speaker

New York Times' Connections: A Case Study on NLP in Word Games

The Tree Center

1 talk · 1 speaker

LLMs for the working programmer. Become a 10x programming centaur today!

TheBrain.pro

1 talk · 1 speaker

Books reimagined: AI to create new experiences for things you know — Łukasz Gandecki, TheBrain.pro

Theory Ventures

1 talk · 1 speaker

Where AI is superhuman: The right jobs to automate with LLMs

Theta Software

1 talk · 1 speaker

Rethinking Environments for Long Horizon Work

Thinking Machines

1 talk · 1 speaker

Frontier Feud

thryv.com

1 talk · 1 speaker

The Demo I Wish I'd Had: OpenAI's Agents SDK... serverless!

Tinder

1 talk · 1 speaker

AI Frontiers in Trust and Safety: Combatting Multifaceted Harm on Tinder at Scale

TNG Technology Consulting

1 talk · 1 speaker

Running a Chess YouTube Channel Entirely by AI — Stephan Steinfurt, TNG Technology Consulting

Tola Capital

1 talk · 1 speaker

E-Values: Evaluating the Values of AI

Towards AI Inc

1 talk · 1 speaker

Build Your Own Deep Research Agent + Technical Writer

Trainline

1 talk · 2 speakers

Shipping complex AI applications | Braintrust & Trainline

Trajectory

1 talk · 1 speaker

Scaling up Continual Learning

Traversal

1 talk · 2 speakers

Production software keeps breaking, and it will only get worse. Here's how Traversal is fixing it.

Trigger.dev

1 talk · 1 speaker

Two Roads to Durable Agents: Replay vs. Snapshot — Eric Allam, Co-founder, Trigger.dev

Trunk Tools

1 talk · 1 speaker

Trunk Tools Launch: Disrupting the $15 Trillion Construction Industry with Autonomous Agents

turbopuffer

1 talk · 1 speaker

Benchmarking semantic code retrieval on Claude Code

Tusk

1 talk · 1 speaker

Designing AI to Scale Human Thought

TwelveLabs

1 talk · 1 speaker

Video Has No Memory. Here's How We Built One.

TypeSafe AI

1 talk · 1 speaker

What's next after RLHF?

Ufonia

1 talk · 1 speaker

Shipping AI to a Million Patients Without an A/B Test

Unsloth AI

1 talk · 1 speaker

Advanced: Reinforcement Learning, Kernels, Reasoning, Quantization & Agents — Daniel Han

Untapped Capital

1 talk · 1 speaker

Active Graph Agent Runtime (BabyAGI 4)

uRun

1 talk · 1 speaker

Generative Video at the Speed of Light

Varick Agents

1 talk · 1 speaker

AI tools for Forward Deployed Engineering

Vellum

1 talk · 1 speaker

AI Agents, Meet Test Driven Development

Vibe Kanban

1 talk · 1 speaker

Software Engineering Is Becoming Plan and Review

Viktor

1 talk · 1 speaker

Viktor — AI Coworker That Lives in Slack

VisualLabs

1 talk · 1 speaker

You Can't Prompt the Room: The Last Skill AI Won't Replace

W&B from CoreWeave

1 talk · 1 speaker

The Z/L Continuum: Should AI Engineers Still Read Code?

Wandero AI

1 talk · 1 speaker

The Missing Layer After Launch

Wasp

1 talk · 1 speaker

GPT Web App Generator - 10,000 apps created in a month: Matija Sosic

Watershed Technology Inc.

1 talk · 1 speaker

Respect The Process

Waymo

1 talk · 1 speaker

Waymo's EMMA: Teaching Cars to Think - Jyh-Jing Hwang, Waymo

Waypoint AI

1 talk · 1 speaker

Cognitive Exhaust Fumes, or: Read-Only AI Is Underrated — Šimon Podhajský, Head of AI, Waypoint

Weco AI

1 talk · 1 speaker

How Autoresearch Is Changing ML Research — Zhengyao Jiang, Weco AI

WEKA

1 talk · 2 speakers

Context Platform Engineering to Reduce Token Anxiety — Val Bercovici and Callan Fox, WEKA

WhyHow.AI

1 talk · 1 speaker

Knowledge Graphs in Litigation Agents — Tom Smoker, WhyHow.AI

Witan Labs

1 talk · 1 speaker

Teaching Coding Agents to do Spreadsheets

Workday

1 talk · 1 speaker

Build Dynamic Products, and Stop the AI Sideshow

You.com / Recursive Superintelligence

1 talk · 1 speaker

First Steps Toward Automated AI Research

Z.ai

1 talk · 1 speaker

Z.ai GLM-4.6: What We Learned From 100 Million Open Source Downloads — Yuxuan Zhang, Z.ai

ZenML

1 talk · 1 speaker

Your Agents Need a Save Button

Zep AI

1 talk · 1 speaker

Citation Needed: Provenance for LLM-Built Knowledge Graphs

Zeta Labs

1 talk · 2 speakers

Which Jobs Can Be Replaced Today

ZS

1 talk · 2 speakers

Why We Killed Our Multi-Agent Pipeline — Subbiah Sethuraman and Abhilash Asokan, ZS Associates