Looking to learn about a new topic or skill? Look no further! Gemini's Guided Learning feature acts as your own personal tutor, teaching you about a particular subject through an engaging back and forth conversation. This AI tool helps users develop their knowledge and skills on a wide variety of topics, acting as a patient mentor, breaking down complex topics step-by-step. This session will take place on 2/24 at 11 AM. Please register using the link below!
https://stonybrookuniversity.co1.qualtrics.com/jfe/form/SV_a9PVlBw0E1Bal1A?

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes three short talks on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

Speakers

Kriti Chopra, Computing & Data Sciences (CDS)
Thomas Flynn, Computing & Data Sciences (CDS)
Wenjie Liao, Chemistry Division

Tuesday, January 7, 2025, 12:00 pm -- CDS, Bldg. 725, Training Room

Join ZoomGov Meeting: https://bnl.zoomgov.com/j/1615289117?pwd=Hqkbj9itxWrFnkhZ8rQXHPInO2gxdF.1

Meeting ID: 161 528 9117
Passcode: 991382

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes one short talk on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

AI and Edge Processing Co-Design for Radiation Detectors

Abstract: Artificial Intelligence (AI) offers exciting new opportunities for enhancing the performance of radiation detectors, ultimately leading to improved physics outcomes. Furthermore, with the explosive growth in data rates being seen by next-generation radiation detectors, deployment of AI algorithms at the edge by embedding intelligence within or near the detector front-end can be transformative. Such integration enables real-time data filtering, noise suppression, feature extraction, and adaptive control, while reducing downstream bandwidth and power consumption. This talk will cover three efforts that bring AI to the forefront of detector technology. First, we demonstrate how AI-based algorithms can be used for position reconstruction in virtual Frisch-grid (VFG) detectors by compensating for charge transport distortions and detector non- uniformities, leading to significantly enhanced fidelity in imaging of gamma-ray interactions. Second, we present a smart readout application specific integrated circuit (ASIC) that combines digital signal processing with co-designed artificial neural networks to enable on-chip regression and classification of detector signals, while meeting stringent constraints on accuracy, speed, and area. Finally, we introduce our recent efforts related to the development of electro-photonic processing architectures that integrate CMOS electronics and silicon photonics for near-sensor AI acceleration. These architectures aim to leverage cross-disciplinary co-design from algorithms to hardware, to achieve low latency and energy-efficient processing of detector data.

Biography: Dr. Prashansa Mukim is an early-career researcher in the Instrumentation Department at BNL, where she works on the design of front-end electronics for extreme environments and the development of co-design methodologies for novel processing modalities and beyond-CMOS technologies. Prior to joining BNL, she was a post-doctoral researcher at the National Institute of Standards and Technology (NIST) in Maryland, where she focused on characterizing the properties of CMOS circuits at cryogenic temperatures and applications of spintronic devices for neuromorphic computing. She received her Ph.D. in Electrical and Computer Engineering from the University of California, Santa Barbara, in 2021.

Location: CDS, Bldg. 725, Training Room

Join ZoomGov Meeting: https://bnl.zoomgov.com/j/1608585935?pwd=UemgEkqijfNf3vIJIGuOa2MdjsunaT.1

Meeting ID: 160 858 5935
Passcode: 076033

Abstract: Generative visual models like Stable Diffusion and Sora generate photorealistic images and videos that are nearly indistinguishable from real ones to a naive observer. However, their grasp of the physical world remains an open question: Do they understand 3D geometry, light, and object interactions, or are they mere pixel parrots of their training data? Through systematic probing, I will demonstrate that these models surprisingly learn fundamental scene properties--intrinsic images such as surface normals, depth, albedo, and shading (à la Barrow & Tenenbaum, 1978)--without explicit supervision, which enables applications like image relighting. But I will also show that this knowledge is insufficient. Careful analysis reveals unexpected failures: inconsistent shadows, multiple vanishing points, and scenes that defy basic physics. All these findings suggest these models excel at local texture synthesis but struggle with global reasoning: a crucial gap between imitation and true understanding. I will then conclude by outlining a path toward generative world models that emulate global and counterfactual reasoning, causality, and physics.

Bio: Anand Bhattad is a Research Assistant Professor at the Toyota Technological Institute at Chicago. He earned his PhD from the University of Illinois Urbana-Champaign in 2024 under the mentorship of David Forsyth. His research interests lie at the intersection of computer vision and computer graphics, with a current focus on understanding the knowledge encoded in generative models. Anand has received Outstanding Reviewer honors at ICCV 2023 and CVPR 2021, and his CVPR 2022 paper was nominated for a Best Paper Award. He actively contributes to the research community by leading workshops at CVPR and ECCV, including Scholars and Big Models: How Can Academics Adapt? (CVPR 2023), CV 20/20: A Retrospective Vision (CVPR 2024), Knowledge in Generative Models (ECCV 2024), and How to Stand Out in the Crowd? (CVPR 2025). For more details, visit https://anandbhattad.github.io/


  • CEWIT's 6th annual hackathon sponsored by Major League Hacking, Hack@CEWIT2022, is taking place virtually on February 18-20, 2022. This year's theme is Hacking Into the Metaverse and will focus on NFT's, Blockchain, Crypto, and the Metaverse. To find out more about the event, mentoring, sponsoring, or to register, visit:

  • https://www.cewit.org/programs/events/hack.php


The AI Innovation Institute cordially invites you to the Summer Symposium this Friday, July 31st, from 10:00 AM to 12:00 PM in the Stony Brook Union Ballroom.



Poster Presentations: 10:00 - 11:30 am - Union Ballroom
Closing remarks: (11:25) Dr. Carl Lejuez, Executive Vice President & Provost
Group Photos (11:30-12)

Featuring undergraduate researchers/participants of the following:
  • AI Innovation & Diffusion Research Experience for Undergraduates (REU)
  • Explorations in STEM
  • Frances Velay Fellowship Program
  • Oncology Research and Clinical Learning Experience (ORACLE)
  • SUNY EOP
  • SUNY SOAR
  • URECA REACT/Research Entry Accelerator for College Transfers
  • URECA Summer

Join us in celebrating the hard work and accomplishments of our students.

Abstract: Pre-trained diffusion and flow matching models have made visual generation remarkably powerful, enabling high-fidelity synthesis of images and videos from natural language prompts. However, their behavior is still largely dictated by the pre-training data distribution and likelihood objective, which do not directly encode downstream desiderata such as fine-grained semantic alignment, controllability, or realism. This gap motivates post-training: starting from a base generator and further optimizing it with additional supervision signals derived from human or reward model preferences.This work presents post-training for visual generative models through two complementary case studies. First, Hummingbird addresses the problem of fine-grained contextual alignment in image-text-to-image generation. We introduce a multimodal context evaluator that scores the consistency between rich contextual descriptions and generated images, capturing fine-grained alignment beyond global CLIP similarity. By directly backpropagating these differentiable rewards through the diffusion sampler, Hummingbird substantially improves semantic faithfulness while preserving high visual quality.
Second, PISCES tackles post-training for text-to-video generation, where alignment is inherently semantic-spatio-temporal. We show that naive VLM-based rewards suffer from distributional mismatch and token-level misalignment, leading to reward hacking and suboptimal optimization. PISCES introduces a bi-objective, Optimal Transport (OT)-aligned reward module: distributional OT using Neural Optimal Transport to align text and video embedding distributions, and discrete, partial OT over a spatio-temporal cost matrix to capture semantic alignment at the token level. These rewards are integrated into both direct backpropagation and GRPO-style optimization to post-train state-of-the-art text-to-video generators. Together, Hummingbird and PISCES provide a unified view of how carefully designed visual reward models, coupled with OT-based representation alignment, can reliably improve the downstream behavior of pre-trained image and video generators.

Speaker: Minh Quan Le

Location: NCS 220

Zoom: https://stonybrook.zoom.us/j/94798224254?pwd=CFraer25qnpORbJ14aAVHRwaSJOjJM.1
Prof. Eugene A. Feinberg, from the Department of Applied Mathematics and Statistics, presents, Recent Developments in Markov Decision Processes Relevant to AI on April 4 at 4p. The talk discusses recent developments in Markov Decision Processes potentially relevant to artificial intelligence. These developments include complexity estimations for exact and approximate algorithms, decision making with incomplete information and multiple criteria, and continuity properties of optimal values and expectations. Dr. Eugene A. Feinberg is currently Distinguished Professor in the Department of Applied Mathematics and Statistics at Stony Brook University. He is an expert on applied probability, stochastic models of operations research, Markov decision processes, and on industrial applications of operations research and statistics. He has published more than 150 papers and edited the Handbook of Markov Decision Processes. His research has been supported by NSF, DOE, DOD, NYSTAR (New York State Office of Science, Technology, and Academic Research), NYSERDA (New York State Energy Research and Development Authority) and by industry. He is a Fellow of INFORMS (The Institute for Operations Research and Management Sciences) and has received several awards including 2012 IEEE Charles Hirsh Award for developing and implementing smart grid technologies, 2012 IBM Faculty Award, and 2000 Industrial Associates Award from Northrop Grumman. Dr. Feinberg is an Associate Editor for Mathematics of Operations Research and for Applied Mathematics Letters. He is an Area Editor for Operations Research Letters. Refreshments will be provided

Discover how Google Gemini can help you draft lesson plans, generate discussion questions, design activities, and create course materials in a fraction of the time.

You have a workshop to plan, need an engaging activity, and a stack of content to create, but not enough time? Google Gemini, available through your SBU Google account, can help you brainstorm, draft, and refine lesson plans, learning objectives, discussion prompts, rubrics, and more. This session is designed for anyone who is interested in using Gemini to help generate learning activities.

In this session, you will:

  1. Access Google Gemini using your SBU Google account
  2. Write effective prompts for lesson planning and content creation
  3. Generate and refine learning objectives, discussion questions, and activities
  4. Create or improve course materials such as rubrics, instructions, and assessment ideas
  5. Apply best practices for reviewing and adapting AI-generated content

Register here.
CG Group member (and SBU faculty) Chao Chen will speak on Fri, March 12, about the use of topological data analysis in machine learning for image analysis.
Chao has shared some of his research with the CG Group previously, and this will be a great opportunity to learn more about this exciting research area related to computational geometry/topology!

Time: Friday, March 12, 2pm-3pm
Place: Zoom
https://stonybrook.zoom.us/my/profweizhu?pwd=RjVIVXg3YUhudzZZQ3pheHUydTJBUT09



Title: Learning with Topological Information - Image Analysis and Label Noise
Speaker: Prof. Chao Chen (SBU)

Abstract: Modern machine learning faces new challenges. We are
analyzing highly complex data with unknown noise. Topology provides
novel structural information to model such data and noise. In this
talk, we discuss two directions in which we are using topological
information in the learning context. In image analysis, we propose a
topological loss to segment and to generate images with not only
per-pixel accuracy, but also topological accuracy. This is necessary
in analysis of images of fine-scale biomedical structures such as
neurons, vessels, etc.  Extracting these structures with correct
topology is essential for the success of downstream
analysis. Meanwhile, we discuss how to use topological information to
train classifiers robust to label noise. This is important in practice
especially when we are using deep neural networks which tend to
overfit noise. These results have been published in NeurIPS, ECCV,
ICML and ICLR.