Communication-Efficient Heterogeneity-Aware Machine Learning System and Architecture by Xuehai Qian

ABSTRACT: The key success of deep learning is the increasing size of models that can achieve high accuracy. At the same time, it is difficult to train the complex models with large data sets. Therefore, it is crucial to accelerate training with distributed systems and architectures, where communication and heterogeneity are two key challenges. In this talk, I will present two heterogeneity-aware decentralized training protocols without communication bottleneck. Specifically, Hop supports arbitrary iteration gap between workers by novel queue-based synchronization which can tolerate heterogeneity with system techniques. Prague uses randomized communication to tolerate heterogeneity with a new training algorithm based on partial reduce -- an efficient communication primitive. If time permits, I will present the systematic tensor partitioning for training on heterogeneous accelerator arrays (e.g., GPU/TPU). We believe that our principled approaches are crucial for achieving high-performance and efficient distributed training.

BIO: Xuehai Qian is an assistant professor at University of Southern California. His research interests include domain-specific systems and architectures, performance tuning and resource management of cloud systems and parallel computer architectures. He received his PhD from the University of Illinois Urbana Champaign and was a postdoc at UC Berkeley. He is the recipient of W.J Poppelbaum Memorial Award at UIUC, NSF CRII and CAREER Award, and the inaugural ACSIC (American Chinese Scholar In Computing) Rising Star Award.
Spring 2026, Wednesdays 2 to 3:20 pm, NCS 220 and Zoom link to be announced soon.

The seminar will be jointly taught by Prof. Dimitris Samaras (samaras@cs.stonybrook.edu).

The overall purpose of this seminar is to bring together people with interests in Computer Vision theory and techniques and to examine current research issues. This course will be appropriate for people who already took a Computer Vision graduate course or already had research experience in Computer Vision.

To enroll in this course, you must either: (1) be in the Ph.D. program or (2) receive permission from the instructors.

Each seminar will consist of multiple short talks (around 15 minutes) by multiple students. Students can register for 1 credit for CSE656. Registered students must attend and present a minimum of 2 talks. Registered students must attend in person. Up to 3 absences will be excused. Everyone else is welcome to attend.

Please note: Exceptionally, the first meeting on 1/28 will be in NCS 120.

Understand Prompting the crucial part to interface with models

Discover how to prompt effectively by exploring the details behind your AI interactions. This isn't just about basic prompting; it's about understanding how to articulate your ideas clearly. We'll showcase a few prompts and how they work. Discover how giving AI the right details can truly boost your productivity and help you reclaim valuable time in your day.

In this session, you will

  1. Utilize AI models effectively
  2. Understanding different prompts
  3. Find out tips that we use with AI

Register: https://stonybrookuniversity.co1.qualtrics.com/jfe/form/SV_dht1o3rNzlZhHka?source=event+manager&session=0805251000ai

Reception to follow.

Abstract:
In this talk, I will present our journey of developing diverse, adaptive, uncertainty-calibrated AI planning agents that can robustly communicate and collaborate for multi-agent reasoning (on math, commonsense, coding, etc.) as well as for interpretable, controllable multimodal generation (across text, images, videos, audio, layouts, etc.). In the first part, we will discuss improving reasoning via multi-agent discussion among diverse LLMs and structured distillation of these discussion graphs (ReConcile, MAGDi), adaptively learning to balance abstraction, decomposition, refinement, and fast+slow thinking in LLM-agent reasoning (ReGAL, ADaPT, MAgICoRe, System-1.x), as well as confidence calibration in LLMs via speaker-listener pragmatic reasoning and making LLMs better teammates via multi-agent positive-negative persuasion balancing (LACIE, PBT). In the second part, we will discuss interpretable and control-lable multimodal generation via LLM-agents based planning and programming, such as layout-controllable image generation (and evaluation) via visual programming (VPGen+VPEval), consistent multi-scene video generation via LLM-guided planning (VideoDirectorGPT), interactive and composable any-to-any multimodal generation (CoDi, CoDi-2), as well as feedback-driven multi-agent interaction for adaptive environment/data generation via weakness discovery (EnvGen, DataEnvGym).
Bio:
Dr. Mohit Bansal is the John R. & Louise S. Parker Distinguished Professor and the Director of the MURGe-Lab (UNC-NLP Group) in the Computer Science department at UNC Chapel Hill. He received his PhD from UC Berkeley in 2013 and his BTech from IIT Kanpur in 2008. His research expertise is in natural language processing and multimodal machine learning, with a particular focus on multimodal generative models, grounded and embodied semantics, faithful language generation, and interpretable, efficient, and generalizable deep learning.

Imagine machines that can see beyond human limitations--drones locating hidden survivors, cameras predicting structural failures, or medical devices detecting tumors beneath the skin. Traditional vision systems are constrained by the boundaries of human perception, missing vast information present in light interactions. This talk explores the development of advanced vision systems that capture underutilized dimensions of light, model intricate light-scene interactions, and extract hidden 3D information--around corners, beneath surfaces, and at high speeds. By jointly developing novel imaging hardware, efficient rendering models, and physics-based learning algorithms, we aim to transcend conventional vision capabilities--unlocking critical applications in autonomous navigation, structural monitoring, and non-invasive medical imaging.

Speaker Bio:


Akshat Dave is a Postdoctoral Associate at MIT Media Lab in the Camera Culture group working with Prof. Ramesh Raskar. He received his Ph.D. from Rice University ECE Department in 2023 where he was advised by Prof. Ashok Veeraraghavan. His research lies at the intersection of applied optics, computer graphics, and computer vision. His research focuses on developing vision systems that go beyond human perception. His work has been recognized by Rice University's Best Thesis Award, OSA Best Paper Prize, and fellowships by Texas Instruments and Qualcomm.
The Institute for AI-Driven Discovery and Innovation hosts Dr. Mary
Simoni for a talk on her music and its intersection with AI, as part
of the Music and AI Seminars series.

The event will be held on Thursday, December 10, 2020, at 3:00 PM.

Abstract: Mary Simoni, Dean of Humanities, Arts & Social Sciences at
Rensselaer Polytechnic Institute will discuss her research in the use
of computer algorithms and technology in the composition and
performance of music. The talk will feature compositions inspired by
Augmented Transition Networks (ATNs), employ motion tracking to
control synthesis parameters, and a work in progress that employs
machine learning using training data that juxtaposes classical music
with COVID-19. During this talk, participants will be introduced to
several technologies that support music information retrieval, machine
learning, and algorithmic composition such as jSymbolic, Weka, and
Common Music.

Zoom details below:
https://stonybrook.zoom.us/j/98236706900?pwd=bDFEZFZtaHBWU0cyL0wxK3UrdUpIdz09
Meeting ID: 982 3670 6900
Passcode: 133945  

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes three short talks on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

Tuesday, January 7, 2025, 12:00 pm -- CDS, Bldg. 725, Training Room


Speakers

Sanket Jantre
Tao Zhang
Xi Yu


Join ZoomGov Meeting: https://bnl.zoomgov.com/j/1615289117?pwd=Hqkbj9itxWrFnkhZ8rQXHPInO2gxdF.1

Meeting ID: 161 528 9117
Passcode: 991382

This is Stony Brook's quantum moment. Join us for a spotlight on the core achievements and research excellence of faculty across the Colleges of Arts and Sciences (CAS), and Engineering and Applied Sciences (CEAS) - and their collaborative advancements in quantum science and technology. Learn about the real world impact of their enduring work, their leadership in translating foundational science into entrepreneurial opportunities, and their impetus for making connections to next generation innovation.

Presented by: Catherine Chen, Ph.D., Research Development Associate

Welcome remarks: President Andrea Goldsmith

Panel moderators: Dean David Wrobel, CAS, and Dean Andrew Singer, CEAS

Presentations and panel featuring our faculty:

  • Jennifer Cano, CAS, Physics and Astronomy

  • P. Scott Carney, CEAS, Mechanical Engineering

  • Hyeongrak Chuck Choi, CEAS, Electrical and Computer Engineering

  • Eden Figueroa, CAS, Physics and Astronomy

  • Humanshu Gupta, CEAS, Computer Science

  • Angela Kelly, CAS, Physics and Astronomy

Location: Theatre at the Charles B. Wang Center, Stony Brook University

Reserve your tickets by March 26!

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes one short talk on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

#1 How to train your Scientific Chatbot by Alexandr Prozorov, Post-Doctoral Research Associate


Abstract: RHIC is closing its 25-year run with ~1 EB of data and decades of hard-won know-how that risk drifting into obscurity. The RHIC Data & Analysis Preservation Plan (DAPP) pilots an AI assistant that lets physicists talk to RHIC in natural language--searching internal notes, code, workflows, and docs, and pointing to runnable, containerized analyses. Built on Retrieval-Augmented Generation(RAG) with a Model Context Protocol orchestration layer, the system indexes heterogeneous, experiment-specific content and enforces role-aware access
for public vs. collaboration-restricted materials. Takeaway: domain-adapted AI can turn a legacy exabyte into reproducible answers, training assets, and new discovery paths.

Biography: Alexandr Prozorov is a postdoc from Czech Technical University in Prague working in STAR experiment. Fascinated by AI

#2 Quantum AI: Atoms, Cavities and Learning by Raman Kumar, Post-Doctoral Research Associate, Instrumentation Department

Abstract: The Instrumentation Department (IO) in the Discovery Technologies directorate at BNL is engaged in exploring various aspects of quantum systems research. One of the main goals of our group's effort is in developing neutral atom-cavity array platforms for remote entanglement generation and distributed quantum processing. This platform promises to herald truly scalable quantum computing systems and open new paradigms for networking and sensing. In this talk, I will explain our group's research and the role AI is playing in unlocking new insights with two examples. The first application of AI is in fabrication process prediction of micro-cavity structures. The second application revolves around role of AI in quantum error detection and correction in modern quantum computing systems.

Biography: Dr. Raman Kumar is a postdoctoral research associate in the IO department at BNL working with Dr./Prof. Sebastian Will (Columbia U.). Kumar obtained his Ph.D. degree in Electrical and Computer Engineering from the University of Illinois Urbana-Champaign. Prior to joining BNL in Nov 2024, Kumar worked as a postdoc at the City College in New York working on topological photonic quantum sensing using NV centers in diamond. Kumar and Will combined have an extremely wide moat and expertise in a variety of different areas which include Ultra cold atoms and molecules, quantum optics, quantum condensed matter, nanofabrication, semiconductor devices and advanced electromagnetics. Their areas of research interest include scalable quantum computing, communications and sensing, all enabled by AI.

Location: CDS, Bldg. 725, Training Room

Join ZoomGov Meeting https://bnl.zoomgov.com/j/1607892208?pwd=MSjxN5btSeToZsQMwEQzCCbBo5h58V.1

Meeting ID: 160 789 2208
Passcode: 753871