BEGIN:VCALENDAR
VERSION:2.0
PRODID:-//Computer Science and Engineering - ECPv6.13.0//NONSGML v1.0//EN
CALSCALE:GREGORIAN
METHOD:PUBLISH
X-ORIGINAL-URL:https://homecse.iitd.ac.in
X-WR-CALDESC:Events for Computer Science and Engineering
REFRESH-INTERVAL;VALUE=DURATION:PT1H
X-Robots-Tag:noindex
X-PUBLISHED-TTL:PT1H
BEGIN:VTIMEZONE
TZID:Asia/Kolkata
BEGIN:STANDARD
TZOFFSETFROM:+0530
TZOFFSETTO:+0530
TZNAME:IST
DTSTART:20250101T000000
END:STANDARD
END:VTIMEZONE
BEGIN:VEVENT
DTSTART;TZID=Asia/Kolkata:20250528T120000
DTEND;TZID=Asia/Kolkata:20250528T130000
DTSTAMP:20260924T034241
CREATED:20250518T062803Z
LAST-MODIFIED:20250526T171039Z
UID:1592-1748433600-1748437200@homecse.iitd.ac.in
SUMMARY:Enhancing Safety and Ethical Alignment in Large Language Models by Rima Hazra
DESCRIPTION:Speaker:  Dr. Rima Hazra \n\n\nAbstract: In this talk\, we explore cutting-edge strategies for enhancing the safety and ethical alignment of large language models (LLMs). The research spans various approaches\, including red teaming and jailbreaking techniques\, which assess and improve model robustness and ethical integrity. We delve into how instruction-centric responses\, when generated by LLMs\, can increase the likelihood of unethical output\, thereby highlighting the vulnerabilities of these AI systems. Through the introduction of frameworks like ‘Safety Arithmetic’ and ‘SafeInfer\,’ we demonstrate methods to mitigate risks by manipulating model parameters and decoding-time behaviors to foster safer interactions. The discussions also emphasize the importance of safety alignment strategies and the challenges posed by integrating new knowledge through model edits\, which can paradoxically destabilize ethical guidelines. This comprehensive examination not only sheds light on the current vulnerabilities of LLMs but also presents a pathway toward more reliable and ethically aligned AI implementations. \n\nBio: Dr. Rima Hazra is a senior postdoc at Eindhoven University of Technology (TU\e)\, Netherlands. Earlier she was a Postdoctoral Researcher at the Singapore University of Technology and Design\, working in the area of AI safety alignment\, natural language processing\, and LLM reasoning. She earned her Ph.D. from the Indian Institute of Technology\, Kharagpur\, where she explored the area of Information retrieval\, NLP and graph learning. With experience in information retrieval\, NLP and graph learning\, Dr. Hazra has published several papers in prestigious CORE A* and A conferences such as AAAI\, ACL\, EMNLP\, NAACL\, ECIR\, ECMLP PKDD and JCDL. She has also received the prestigious Microsoft Academic Partnership Grant (MAPG) and the PaliGemma Academic Program award from Google for her work in AI safety alignment.
URL:https://homecse.iitd.ac.in/event/enhancing-safety-and-ethical-alignment-in-large-language-models-by-dr-rima-hazra/
LOCATION:Bharti 501\, IIT Campus\, Hauz Khas\, New Delhi
CATEGORIES:Seminars
END:VEVENT
END:VCALENDAR