IACS Research Theme: Human Centered Computing Seminar

Abstract: The AI art platform Artbreeder hosts daily remix parties where users build on each other's work, creating transparent evolutionary chains of images from a single seed. This study analyzes 130,882 images from 368 remix parties to identify the drivers of novelty, complexity, and competitive success. The results reveal an interesting tension: while more novel parent images produce more novel and complex children and attract more likes, users paradoxically prefer to remix images that are less novel and complex. At the group level, larger remix parties produce more novelty at the cost of lower complexity. Additionally, images tend to converge towards common thematic attractors (e.g., steampunk scenes, alien architecture, furries) over the course of remix parties. These results provide quantitative insights into collective creativity--the production of novelty by groups of people--a typically opaque aspect of human cultural evolution.

Speaker: Dr. Mason Youngblood

Location: Institute for Advanced Computational Science, Seminar Room
Abstract: In today's digital era, language functions not only as a medium of information transmission but also as a mechanism of persuasion, framing, and control. The proliferation of online platforms has amplified this dual role: while enabling unprecedented access to knowledge, it has also exacerbated challenges such as misinformation, rhetorical manipulation, and cultural or linguistic disparities in information access. As a result, pragmatic language understanding and information integrity have emerged as central concerns for both computational linguistics and society at large. This research follows how claims are produced, reframed, and contested online through three interconnected threads. First, it models pragmatic deflection in discourse by investigating whataboutism, a rhetorical device that deflects criticism by redirecting discourse, and introduced novel datasets from Twitter (now X) and YouTube. This work underscores how subtle pragmatic maneuvers can erode discourse integrity without relying on outright falsehoods. Second, it advances retrieval and alignment for information integrity in health and news communication. These systems trace claims and narratives across genres (e.g., social posts and news reports) and languages (Chinese and English), linking social posts with journalistic reporting and aligning Chinese news with English biomedical evidence. By accounting for cultural context, assertions can be linked to reliable evidence and organized for systematic comparison. This work surfaces the risks of missing sources, unverifiable claims, and framing disparities in global health discourse, and demonstrates computational solutions that enhance both the credibility and accessibility of information. Third, the methodological centerpiece is Class Distillation (ClaD), a geometry-aware training paradigm for distilling a small, well-defined target class from a large, heterogeneous background. ClaD couples a distribution-aware contrastive loss (instantiated here in a Mahalanobis form when its assumptions fit the data) with an interpretable decision algorithm tuned for class separation. Evaluated on sarcasm, metaphor, and sexism detection, ClaD delivers strong efficiency and robustness, matching or surpassing larger models while using fewer computational resources, making these pipelines practical by learning reliably from small, sharply defined classes. In sum, this research presents an integrated account of language understanding in the digital age. It exposes how integrity falters through pragmatic deflection, cross-genre drift, and cross-lingual misalignment, and translates these insights to move pragmatic language understanding to systems for evidence retrieval, alignment, and verification; and it sheds light on where and how integrity is threatened, and delivers methods that leverage pragmatic language use.

Speaker: Chenlu Wang

Location: (Old) Computer Science Building, Room 2311
The INS (International Neuroethics Society) AI and Consciousness Affinity Group is hosting a talk titled Bringing Trustworthiness in Generative AI and Agentic AI Using Thought Knowledge Graphs featuring speaker Manas Gaur, a computer scientist at UMBC.
The talk will examine the interplay between Thought Knowledge Graphs (TKGs) and how they can form more trustworthy and reasoning-based responses in AI. They will also discuss introducing novel methods on implementing TKGs and their overall impact on creating more trustworthy AI systems.
The talk will be held online via Zoom on Monday, December 2 at 1:00pm (EST).
Register to attend.

The AI Community will be hosting our very first Datathon๐Ÿ’ก๐Ÿ“Š

Ready to turn data into groundbreaking insights? ๐Ÿง 

Compete in our Datathon, where you'll analyze real-world data ๐Ÿ“ˆ and share innovate solutions in these tracks:

๐Ÿซ Student Life

๐ŸŒฑ Environment & Sustainability

๐Ÿ’‰ Health & Wellness

๐Ÿ’ฐ Finance & Economics

Whether you're a data pro or just starting out, this is your chance to network, learn, and win exciting prizes! ๐Ÿ†๐ŸŽ‰ Bring your creativity ๐Ÿงฉ collaborate with fellow students ๐Ÿง‘โ€๐Ÿคโ€๐Ÿง‘ and gain hands-on experience showcasing your analytical skills ๐Ÿ’ป

Submissions will be judged by professors ๐Ÿง‘โ€๐Ÿซ so take this chance to impress them!

There will be free food โ˜• and games ๐ŸŽฒ to fuel your brain and imagination! Don't miss out--register now and unleash the power of data! ๐Ÿ”ฅโœจ

Registration Form: https://forms.gle/6XYMfmhyAByzFpxz5

Time: Friday (4/4) 10:30am - 5pm โฐ

Location: Bauman Center ๐Ÿ“

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes one short talk on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Abstract: Designing custom proteins could revolutionize medicine and materials, but it remains an immense scientific challenge. Our work uses large-scale AI foundation models to generate novel proteins tailored to bind specific small molecules. Each AI-generated design is passed through a rigorous, multi-stage validation pipeline to ensure it is biophysically realistic. A key innovation is fine-tuning our model with data from molecular dynamics (MD) simulations, exposing it to the conformational dynamics and energetics of protein-ligand binding. This physics-aware training results in novel protein designs with enhanced stability and more effective binding capabilities.

Bio: Xin Dai is an Assistant Computational Scientist in the Artificial Intelligence Department of the CDS. His work centers on AI for Science with a strong focus on computational biology. He earned his PhD in Physics from Tsinghua University.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

Location: CDS, Bldg. 725, Training Room

Join Zoom Meeting: https://bnl.zoomgov.com/j/1604383624?pwd=ffQ5cUPNxTI7nzClKQO6cnsNbhF9Vf.1

Meeting ID: 160 438 3624
Passcode: 558449

Title: Cultural Biases, World Languages, and User Privacy in Large Language Models
Abstract: In this talk, I will highlight three key aspects of large language models: (1) cultural bias in LLMs and pre-training data, (2) decoding algorithm for low-resource languages, and (3) human-centered design for real-world applications.

The first part focuses on systematically assessing LLMs' favoritism towards Western culture. We take an entity-centric approach to measure the cultural biases among LLMs (e.g., GPT-4, Aya, and mT5) through natural prompts, story generation, sentiment analysis, and named entity tasks. One interesting finding is that a potential cause of cultural biases in LLMs is the extensive use and upsampling of Wikipedia data during the pre-training of almost all LLMs. The second part will introduce a constrained decoding algorithm that can facilitate the generation of high-quality synthetic training data for fine-grained prediction tasks (e.g., named entity recognition, event extraction). This approach outperforms GPT-4 on many non-English languages, particularly low-resource African languages. Lastly, I will showcase an LLM-powered privacy preservation tool designed to safeguard users against the disclosure of personal information. I will share findings from an HCI user study that involves real Reddit users utilizing our tool, which in turn informs our ongoing efforts to improve the design of AI models.
Bio:

Wei Xu is an Associate Professor in the College of Computing and Machine Learning Center at the Georgia Institute of Technology, where she is the director of the NLP X Lab. Her research interests are in natural language processing and machine learning, with a focus on Generative AI, robustness and fairness of large language models, multilingual LLMs, as well as AI for science, education, accessibility, and privacy research. She is a recipient of the NSF CAREER Award, Google Academic Research Award, CrowdFlower AI for Everyone Award, Best Paper Awards and Honorable Mentions at COLING'18, ACL'23, ACL'24. She also received research funds from DARPA and IARPA. She is currently an executive board member of NAACL. Join Zoom Meeting https://stonybrook.zoom.us/j/98855994362?pwd=F2qnpwL85fhCBHAEW9ZBpXihfwGHsj.1 (ID: 98855994362, passcode: 172797) Join by phone (US) +1 646-876-9923 (passcode: 172797) Joining instructions: https://www.google.com/url?q=https://applications.zoom.us/addon/invitation/detail?meetingUuid%3DuDJcUTvyQueZkCaUSAwFlg%253D%253D%26signature%3Da3d49e0f7f2e74e7130f7308c74bd85ba7b99587b98ba2e34238bb657ca51a09%26v%3D1&sa=D&source=calendar&usg=AOvVaw2jTn5cjfRG8vXU8KHHlU2Y Meeting host: H.Andrew.Schwartz@stonybrook.edu

Join Zoom Meeting:
https://stonybrook.zoom.us/j/98855994362?pwd=F2qnpwL85fhCBHAEW9ZBpXihfwGHsj.1
Hosted by the College of Business faculty and staff, this virtual session will explore how our MS in Business Analytics & Intelligence program can enable you to become a future leader with AI expertise. This session will also cover the program's benefits, curriculum, cost, and admission requirements.

Register here.
Presented by Stony Brook University Department of Biomedical Informatics and Long Island Network for Clinical and Translational Science (LINCATS).

The seminar aims to empower participants with the knowledge and skills necessary to harness AI effectively in clinical practice and research. It will equip attendees with practical insights, case studies, and interactive discussions led by experts in both AI and medicine, fostering a collaborative environment where attendee can explore how to overcome barriers and maximize the potential of AI in transforming modern healthcare delivery.

All Stony Brook Audiences Welcome.
Please note: This exciting event is open to all Stony Brook Faculty/Staff/Students. While the overarching theme for this event is the application of AI in medicine, the event is designed to bridge the professional practice gap that exists between cutting-edge AI research and its practical implementation in clinical settings, While AI holds immense promise for transforming healthcare delivery, many physicians and researchers lack the foundational knowledge and practical skills needed to effectively integrate AI into their daily practices.

THIS CONFERENCE IS FOR STONY BROOK UNIVERSITY & HOSPITAL FACULTY/STAFF & STUDENTS ONLY.


Registration link: https://cme.stonybrookmedicine.edu/continuing-medical-education/conferences/235/bench-to-bedside-understanding-the-practical-application-of-ai-in-medicine-2024/10/17/2024

FOR QUESTIONS
joseph.cesaria@stonybrookmedicine.edu
mary.saltz@stonybookmedicine.edu