AI Engineer World's Fair 202433:39Mastering LLM Inference Optimization: From Theory to Cost-Effective DeploymentMark MoyouInfrastructure and deployment · Evals · Architecture