BEGIN:VCALENDAR
VERSION:2.0
PRODID:-//pretalx//pretalx.com//pyconde-pydata-2025//speaker//MUAVLE
BEGIN:VTIMEZONE
TZID:CET
BEGIN:STANDARD
DTSTART:20001029T040000
RRULE:FREQ=YEARLY;BYDAY=-1SU;BYMONTH=10
TZNAME:CET
TZOFFSETFROM:+0200
TZOFFSETTO:+0100
END:STANDARD
BEGIN:DAYLIGHT
DTSTART:20000326T030000
RRULE:FREQ=YEARLY;BYDAY=-1SU;BYMONTH=3
TZNAME:CEST
TZOFFSETFROM:+0100
TZOFFSETTO:+0200
END:DAYLIGHT
END:VTIMEZONE
BEGIN:VEVENT
UID:pretalx-pyconde-pydata-2025-P8GUWG@pretalx.com
DTSTART;TZID=CET:20250425T132000
DTEND;TZID=CET:20250425T135000
DESCRIPTION:Many LLM benchmarks focus on reasoning and coding tasks. These 
 are exciting tasks! But the majority of LLM usage is still in writing and 
 editing related tasks\, and there's a surprising lack of benchmarks on the
 se. \n\nIn this talk you'll learn what it took to create a writing benchma
 rk\, and which model performs best!
DTSTAMP:20260719T220956Z
LOCATION:Platinum3
SUMMARY:Is your LLM any good at writing? Benchmarking on creative writing a
 nd editing tasks - Azamat Omuraliev
URL:https://pretalx.com/pyconde-pydata-2025/talk/P8GUWG/
END:VEVENT
END:VCALENDAR
