BEGIN:VCALENDAR
VERSION:2.0
PRODID:-//pretalx//pretalx.com//pydata-amsterdam2026//talk//R3KMAB
BEGIN:VTIMEZONE
TZID:Europe/Amsterdam
BEGIN:DAYLIGHT
DTSTART:20250912T000000
TZNAME:CEST
TZOFFSETFROM:+0200
TZOFFSETTO:+0200
END:DAYLIGHT
BEGIN:STANDARD
DTSTART:20251026T030000
RDATE:20261025T030000
TZNAME:CET
TZOFFSETFROM:+0200
TZOFFSETTO:+0100
END:STANDARD
BEGIN:DAYLIGHT
DTSTART:20260329T030000
RDATE:20270328T030000
TZNAME:CEST
TZOFFSETFROM:+0100
TZOFFSETTO:+0200
END:DAYLIGHT
END:VTIMEZONE
BEGIN:VEVENT
SUMMARY:Do you know how well your model is doing? Evaluate your LLMs - Che
 uk Ting Ho
DTSTART;TZID=Europe/Amsterdam:20260912T130000
DTEND;TZID=Europe/Amsterdam:20260912T140000
DTSTAMP:20260911T154339Z
UID:pretalx-pydata-amsterdam2026-R3KMAB@pretalx.com
DESCRIPTION:Large Language Models (LLMs) are becoming central to modern ap
 plications\, yet effectively\nevaluating their performance remains a signi
 ficant challenge. How do you objectively compare different models\, benchm
 ark the impact of fine-tuning\, or ensure your LLM responses adhere to saf
 ety guidelines (guard-railing)? \nThis hands-on workshop addresses these c
 ritical questions.
LOCATION:Room B
URL:https://pretalx.com/pydata-amsterdam2026/talk/R3KMAB/
END:VEVENT
END:VCALENDAR
