
From weeks to a day: how we made LLM evaluation fast enough to iterate on
Baharak Saberidokhtmedium.com

Baharak Saberidokhtmedium.com
Beyond the model: Engineering AI infra with scientific judgementAirbnbEngmedium.com
Project Lighthouse — Part 3: Introducing project-lighthouse-anonymizeAdam Bloomstonmedium.com
How we knew COVID was over (and what our models had to unlearn)Harrison Katzmedium.com
Flexible Authentication: Reimagining authentication for millions of users at AirbnbJose Santosmedium.com
Eval-driven development: Lessons from evaluating GenAI at scaleRohit Girmemedium.com