Speakers1 - 5 of 10
Open a speaker to read their bio and see everything they're presenting.
Sessions (1)
- Evaluation harnesses for production LLMs
Monday, October 12: 10:00 AM - 10:45 AM · Workshop Room
AI Engineering
- Evaluation harnesses for production LLMs
Hana works on calibration and abstention — teaching models to recognise the edge of their own competence. She publishes regularly and reviews for three conferences.
Sessions (1)
- Evaluation harnesses for production LLMs
Monday, October 12: 10:00 AM - 10:45 AM · Workshop Room
AI Engineering
- Evaluation harnesses for production LLMs
Sofia was the first engineer at Vantage Labs and now leads their agent runtime. She is unreasonably interested in what happens when a tool call fails halfway through.
Sessions (1)
- Building reliable agents: a practitioner's playbook
Monday, October 12: 10:00 AM - 11:00 AM · Main Stage
AI Engineering
- Building reliable agents: a practitioner's playbook
Ava leads the evaluation platform at Lumen AI, where she is responsible for making sure model changes ship without quietly regressing customer-facing quality. She spent six years in search relevance before moving to LLM systems.
Sessions (2)
- Building reliable agents: a practitioner's playbook
Monday, October 12: 10:00 AM - 11:00 AM · Main Stage
AI Engineering - Closing panel: what we got wrong about agents
Tuesday, October 13: 04:00 PM - 05:00 PM · Main Stage
Product
- Building reliable agents: a practitioner's playbook
Priya runs the platform organisation at Cobalt Systems, covering everything from developer experience to the inference fleet. She has been a programme chair three times and still enjoys it.
Sessions (1)
- Opening keynote: the year AI engineering grew up
Monday, October 12: 09:00 AM - 09:45 AM · Main Stage
AI Engineering
- Opening keynote: the year AI engineering grew up