Title: Cultural Biases, World Languages, and User Privacy in Large Language Models
Abstract: In this talk, I will highlight three key aspects of large language models: (1) cultural bias in LLMs and pre-training data, (2) decoding algorithm for low-resource languages, and (3) human-centered design for real-world applications.

The first part focuses on systematically assessing LLMs' favoritism towards Western culture. We take an entity-centric approach to measure the cultural biases among LLMs (e.g., GPT-4, Aya, and mT5) through natural prompts, story generation, sentiment analysis, and named entity tasks. One interesting finding is that a potential cause of cultural biases in LLMs is the extensive use and upsampling of Wikipedia data during the pre-training of almost all LLMs. The second part will introduce a constrained decoding algorithm that can facilitate the generation of high-quality synthetic training data for fine-grained prediction tasks (e.g., named entity recognition, event extraction). This approach outperforms GPT-4 on many non-English languages, particularly low-resource African languages. Lastly, I will showcase an LLM-powered privacy preservation tool designed to safeguard users against the disclosure of personal information. I will share findings from an HCI user study that involves real Reddit users utilizing our tool, which in turn informs our ongoing efforts to improve the design of AI models.
Bio:

Wei Xu is an Associate Professor in the College of Computing and Machine Learning Center at the Georgia Institute of Technology, where she is the director of the NLP X Lab. Her research interests are in natural language processing and machine learning, with a focus on Generative AI, robustness and fairness of large language models, multilingual LLMs, as well as AI for science, education, accessibility, and privacy research. She is a recipient of the NSF CAREER Award, Google Academic Research Award, CrowdFlower AI for Everyone Award, Best Paper Awards and Honorable Mentions at COLING'18, ACL'23, ACL'24. She also received research funds from DARPA and IARPA. She is currently an executive board member of NAACL. Join Zoom Meeting https://stonybrook.zoom.us/j/98855994362?pwd=F2qnpwL85fhCBHAEW9ZBpXihfwGHsj.1 (ID: 98855994362, passcode: 172797) Join by phone (US) +1 646-876-9923 (passcode: 172797) Joining instructions: https://www.google.com/url?q=https://applications.zoom.us/addon/invitation/detail?meetingUuid%3DuDJcUTvyQueZkCaUSAwFlg%253D%253D%26signature%3Da3d49e0f7f2e74e7130f7308c74bd85ba7b99587b98ba2e34238bb657ca51a09%26v%3D1&sa=D&source=calendar&usg=AOvVaw2jTn5cjfRG8vXU8KHHlU2Y Meeting host: H.Andrew.Schwartz@stonybrook.edu

Join Zoom Meeting:
https://stonybrook.zoom.us/j/98855994362?pwd=F2qnpwL85fhCBHAEW9ZBpXihfwGHsj.1

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes three short talks on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

Tuesday, January 7, 2025, 12:00 pm -- CDS, Bldg. 725, Training Room

Speakers

Jianda Chen, EBNN - Improving the stability and accuracy of PDE-ML hybrid AGCMs

Boyang Li, CDS - Accelerating Materials Discovery using Machine Learning

Jaehye on Do, NPP Isotopes - Using LLMs for Isotopes Research and Production

Join ZoomGov Meeting: https://bnl.zoomgov.com/j/1615289117?pwd=Hqkbj9itxWrFnkhZ8rQXHPInO2gxdF.1

Meeting ID: 161 528 9117
Passcode: 991382

I will be holding an informal 2-week short optimization course, to try
to cover a few important proofs in the field. The goal will be depth
over breadth, with focus on:

 - convergence proofs for gradient descent and stochastic gradient descent
 - energy functions and continuous time optimization
 - estimate sequences and Nesterov acceleration

and, time permitting, additional topics like variance reduction,
quasi-Newton methods, and Frank-Wolfe methods. If we go super fast, we
can spend a few days at the end brainstorming interesting research
project ideas.

Details: NCS 220 6:15pm-7:45pm, Monday-Friday, Feb 7-Feb 18.

In person only, since I plan to use the whiteboard (but may be recorded)

More details will be uploaded here (notes, specific schedule):
https://sites.google.com/view/optimization-short-course/home

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes one short talk on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

Embodied Intelligence at Scientific User Facilities

Abstract: This presentation explores the active work integrating artificial intelligence and robotics at the National Synchrotron Light Source II, and a perspective for the future. Through various case studies, we highlight the optimization of operations, improved experimental outcomes, and the orchestration of distributed multimodal experiments. This ongoing development includes collaborators from across the light and neutron sources in the DOE complex. We will elaborate on the open-source Bluesky project, and its capabilities to support adaptive and autonomous experiments. Additionally, we will discuss how Bluesky can be integrated with open-source robotic control software to unlock new flexible automation for autonomous scientific research, which scales to new experiments and continues to leverage human ingenuity.

Biography: Dr. Phillip M. Maffettone is an Associate Computational Scientist in the Data Science and Systems Integration Division at NSLS-II. His research focuses on accelerating scientific discovery at user facilities through the integration of robotics, artificial intelligence (AI), and advanced experiment orchestration systems. He leads the N3XTware project, constructing the software architecture for the next 12 beamlines to be built at NSLS-II. Prior to this he built the brain on the world's first mobile robotic scientist at the University of Liverpool, and later spearheaded the machine learning platform for a biotechnology start-up, BigHat Biosciences. He holds a DPhil in Inorganic Chemistry from the University of Oxford and a B.S. in Chemical Engineering from the University at Buffalo.

Location: CDS, Bldg. 725, Training Room

Link: https://bnl.zoomgov.com/j/16049713 31?pwd=nc5CV3cOFrdYxordFieP W07tIDmwYb.1

Meeting ID: 160 497 1331
Passcode: 289875

Join us at the Center for Excellence in Learning and Teaching (CELT) for an interactive Zoom workshop on Generative AI designed for faculty and staff interested in enhancing teaching and assessment practices, increasing student engagement, and navigating the rapidly evolving landscape of AI tools. Participants will be introduced to common AI tools, explore potential instructional uses, and discuss key considerations such as academic integrity, transparency, and equity.

Register now: https://stonybrook.zoom.us/meeting/register/6js1eP64T1ys8tyU57EJ7Q#/registration
Presented by the Stony Brook University Hospital Institutional Ethics Committee and co-sponsored by the Center for Medical Humanities, Compassionate Care and Bioethics, Stony Brook University

9th Annual Medical Ethics Symposium

Friday, August 7, 2026: MART Auditorium live and via Webinar

Artificial intelligence (AI) is transforming healthcare by improving diagnosis, streamlining administrative tasks, supporting personalized treatment plans, and aiding medical research. However, the growing use of AI in medicine also raises important ethical concerns. While AI has the potential to enhance patient care and efficiency, healthcare providers and patients must approach these technologies with caution and critical oversight.

This event invites professionals and students from all healthcare, legal, and other associated disciplines to bring their work, challenges, and solutions to the table at 'Trust but Verify: Ethical Challenges of AI in Healthcare.'

Agenda:

8am Registration/Light Breakfast
Provided by The Center for Medical Humanities, Compassionate Care and Bioethics

8:30am Welcome
Jean Mueller, MPS, BS, RN, CPHQ Ethics Symposium Coordinator

Opening Remarks
Carolyn Santora, MS, RN, NEA-BC, CPHQ, Chief Nursing Officer; Chief Nursing Officer; Chief of Regulatory Affairs, Patient Safety and Ethics, Stony Brook University Hospital; Chair/Institutional Ethics Committee, Stony Brook Medicine

KEYNOTE
9- 10:30am Trust, Communication, and Consent in AI-Mediated Healthcare
Kellie Owens, PhD
Assistant Professor, Medical Ethics, Department of Population Health, NYU Grossman School of Medicine

KEYNOTE
10:45- 12:15pm AI That Measures, AI that Speaks: A Decade of Surgical AI and the Problem of Verification
Alexander Winkler- Schwartz, MDCM, PhD, FRCSC, FAANS
Neurosurgeon, Neuroscientist, AI and Surgical Education Researcher
Assistant Professor, Adult and Pediatric Neurosurgery, Stony Brook University

12:30- 1:15pm Lunch
Provided by the Stony Brook University Hospital Institutional Ethics Committee

Panel Discussion and Interactive Ethics Case Presentations
1:30- 3pm Healthcare AI: Designed to Assist, Not Resist Human Judgment
Moderator: David N. Hoffman, JD

Distinguished Panelists
Leah Gancz, MD
Julie Luengas, DNP, MBA, RN, NI-BC, FHIMSS
Tauhid Mahmud, MD, MPH
Pons Materum III, MD
Neil J. Patel, MD, MBA, MS
Carolyn Santora, RN, NEA-BC, CPHQ
Caitlyn Tabor, JD, MBE
Mathew Tharakan, MD, MBA
Alexander Winkler-Schwartz, MDCM, PhD, FRCSC, FAANS
Zhi Wu, MD

3- 3:45pm Poster Awards and Rapid-Fire Sessions

Carolyn Santora, MS, RN, NEA-BC, CPHQ
1st, 2nd and 3rd Place Posters and Colleagues Choice Award

3:45- 4pm Summary and Closing Remarks
Carolyn Santora, MS, RN, NEA-BC, CPHQ

Register here.
Join the Center of Excellence in Wireless and Information Technology (CEWIT) and their co-host IEEE-USA for a livestream panel discussion on Generative Artificial Intelligence (Gen AI). In this engaging livestream, we will dive into the technologies that continue to transform what is possible and explore the dynamic intersection of innovation, creativity, ethics, and Gen AI.

CEWIT is joined by Stony Brook University experts who will provide their insights and perspectives on this rapidly changing technology.

Meet the Panel

Laura Lindenfeld, PhD

Executive Director
Alan Alda Center for Communicating Science®
Dean
School of Communication & Journalism
BIO

Margaret Schedel, PhD
Associate Professor
Composition and Computer Music
Co-Founder
Lyrai
BIO

Steven Skiena, PhD

Interim Director
AI Innovation Institute
Distinguished Professor
Computer Science
BIO

Vivian Zhang
CTO/School Director
NYC Data Science Academy
Chief Data Officer
GoDental.ai
BIO


Register here.