BEGIN:VCALENDAR
VERSION:2.0
PRODID:OpenCms 21.0.22
BEGIN:VTIMEZONE
TZID:Europe/Berlin
X-LIC-LOCATION:Europe/Berlin
BEGIN:DAYLIGHT
TZOFFSETFROM:+0100
TZOFFSETTO:+0200
TZNAME:CEST
DTSTART:19700329T020000
RRULE:FREQ=YEARLY;BYDAY=-1SU;BYMONTH=3
END:DAYLIGHT
BEGIN:STANDARD
TZOFFSETFROM:+0200
TZOFFSETTO:+0100
TZNAME:CET
DTSTART:19701025T030000
RRULE:FREQ=YEARLY;BYDAY=-1SU;BYMONTH=10
END:STANDARD
END:VTIMEZONE
BEGIN:VEVENT
DTSTAMP:20241127T123956
UID:53a3a99c-acb4-11ef-9676-000e0c3db68b
SUMMARY:IRIS Colloquium | Analysis of behavior patterns of LLMs
DESCRIPTION:Offensive speech is highly prevalent on online platforms. Being trained on online data, Large Language Models (LLMs) display undesirable behaviors, such as generating harmful text or failing to recognize it. Despite the potential harms from LLMs in such applications, whether LLMs can reliably identify offensive speech and how they behave when they fail are open questions. In this work, we probed sixteen widely used LLMs and showed that most fail to identify (non-)offensive online language. Our experiments reveal undesirable behavior patterns in the context of offensive speech detection, such as erroneous response generation, over-reliance on profanity, and failure to recognize stereotypes.
DTSTART;TZID=Europe/Berlin:20250122T140000
DTEND;TZID=Europe/Berlin:20250122T150000
LOCATION: , U32EGO.131, Universitätsstr. 32,  Campus Vaihingen
URL;VALUE=URI:https://www.iris.uni-stuttgart.de/event/IRIS-Colloquium--Analysis-of-behavior-patterns-of-LLMs/
END:VEVENT
END:VCALENDAR
