Abstract: Visual generation is a fundamental problem in computer vision and graphics, with applications ranging from 3D capture to content creation and image/video synthesis. Despite rapid progress in neural rendering and generative models, efficiency remains a key obstacle in practice: high-quality 3D reconstruction often depends on dense multi-view supervision; scalable 3D synthesis faces heavy optimization, training, and rendering costs; and modern image/video generators incur substantial computation as token grids grow with spatial resolution and temporal length.
This thesis targets efficient visual world modeling by improving sample efficiency in 3D reconstruction, representation efficiency in 3D generation, and computational efficiency in image/video synthesis. First, we improve sample efficiency for neural implicit surface reconstruction under sparse views by integrating multi-view stereo probability volumes as a geometric regularizer, enabling high-quality reconstruction from as few as three input images. Next, we introduce an explicit 3D representation for 3D generation, built from multi-view depth and RGB predictions with 3D Gaussian features, which enables the use of 2D generative priors while enforcing multi-view consistency via epipolar attention. We then address the computational bottleneck of image and video synthesis with importance-based token merging, using importance signals available during generation to preserve critical information while merging redundant tokens. Finally, we propose efficient mixed-resolution diffusion transformers via cross-resolution phase-aligned attention, aiming to improve attention stability under mixed token grids and support high-fidelity mixed-resolution generation.

Speaker: Haoyu Wu

Location: NCS120
Abstract: Machine learning (ML) systems fueled by neural networks have entered our daily lives and led to scientific breakthroughs, but many open questions remain. After a nod toward the question of rigor with ML and recent progress, I'll turn to the theory of neural networks. I will argue that understanding neural networks inevitably leads to ideas from field theory (FT), which was already realized in the simplest case in the 1990s, and I will review some essential FT-for-NN results. I will then propose that the connection might be more general, an NN-FT correspondence of sorts, with neural networks providing a way to define a field theory. I'll end with comments on known results including the origin of interactions and various symmetries, but I will also list some open questions. The apparent non-sequitur in the title will be used as a rhetorical device to explore where we are and where we'd like to go.

https://scgp.stonybrook.edu/calendar/full-calendar
Abstract: At XTX Markets, we view algorithmic trading as one of the most compelling real-world frontiers for deep learning and foundation models. Every day, our systems generate forecasts for tens of thousands of financial instruments and execute over $300B in global trading volume: fully automated, with no discretionary human intervention. This domain combines massive data scale with high noise, adversarial dynamics, and frequent regime shifts, making it both scientifically challenging and commercially impactful. For machine learning researchers, it serves as a rigorous proving ground where advances in time-series modeling, large-scale optimization, representation learning, and foundation models can translate directly into measurable real-world outcomes. This talk will provide a high-level overview of our research agenda, infrastructure, and key open challenges at the intersection of large-scale AI and quantitative finance.

Speaker: Dr. Zhangyang Atlas Wang is the Research Director at XTX Markets, one of the world's leading high-frequency trading firms. He founded and leads the firm's AI Lab in New York City, focused on developing large-scale foundation models for financial time series and market data, powered by XTX's proprietary AI infrastructure. He is currently on leave from his position as the Temple Foundation Endowed Associate Professor at The University of Texas at Austin. His academic research has received numerous awards, and he has mentored a broad network of Ph.D. students and postdoctoral researchers. Many of his alumni now hold tenure-track faculty positions (eight to date) or senior research roles in industry (nineteen and counting). For more information about his group and alumni, please visit: https://www.vita-group.space/team.

Location: NCS 120

Refreshments will be served after the seminar in the first-floor atrium.



This workshop is intended for researchers, practitioners, students, and industry professionals in AI, robotics, machine learning, human-robot interaction, and related fields.

Workshop Overview:

Instead of learning from data alone, an embodied AI system learns through its movements, sensors, and interactions with the environment. This form of active, experience-based learning, informed by ongoing self-evaluation of its own abilities, enables embodied AI systems to adapt on the fly, understand context rather than just commands, and collaborate with humans in more natural and trustworthy ways.

Workshop Goals:

  1. Foster interdisciplinary dialogue across AI, robotics, and cognitive science.
  2. Identify key challenges and future research directions in embodied intelligence.
  3. Examine the role of embodiment in advancing toward AGI.

This workshop is Invitation-only. Please email Dr. IV Ramakrishnan (ram@cs.stonybrook.edu) to attend.

Read the announcement: https://mcusercontent.com/237207911c0fd4c1f78dd8524/files/070dec2e-a2f5-143e-0fe2-c4ebecdb5193/Embodied_AI_Workshop_Invitation_.pdf

Research challenges in using computer vision in robotics systems Abstract The past decade has seen a remarkable increase in the level of performance of computer vision techniques, including with the introduction of effective deep learning techniques. Much of this progress is in the form of rapidly increasing performance on standard, curated datasets. However, translating these results into operational vision systems for robotics applications remains a formidable challenge. This talk with explore some of the fundamental questions at the boundary between computer vision and robotics that need to be addressed. This includes introspection/self-awareness of performance, anytime algorithms for computer vision, multi-hypothesis generation, rapid learning and adaptation. The discussion will be illustrated by examples from autonomous air and ground robots.


Dates: 

Wednesday, March 3, 2021 - 6:00pm to 7:30pm

Location: 

Zoom - contact events@cs.stonybrook.edu for Zoom info.

Event Description: 

Women in Computer Science (WiCS), the Society of Women Engineers (SWE), and the Stony Brook Robotics Team (SBRT) are collaborating to host an event called Inspiring Women in STEM Academia: A Community Dialogue to address the lack of female representation in STEM academia. 
 

All are invited to attend so they may gain a better understanding of the challenges faced by their female colleagues and hear perspectives on how they can offer support in the workplace. Given the shockingly disproportionate number of female professionals in STEM academia, we feel that this event would be extremely beneficial for male faculty to listen to and amplify their voices.

It will begin with a discussion panel consisting of Stony Brook professors and faculty who will provide valuable insight into the issue. From there, we will split into smaller discussion groups where student and faculty attendees will be able to voice their opinions, hear about the thoughts/experiences of others, and participate in an engaging discussion with panelists.

The event will be held on March 3rd from 6:00 - 7:30 PM on Zoom.
 

The following Stony Brook faculty will be panelists:

Dr. Aruna Balasubramanian - Computer Science Professor, WiCS Advisor, WPhD Advisor

Dr. Xinwei Mao - Civil Engineering Assistant Professor

Urszula Zalewski - Director of Experiential Learning, Career Center Advisor (Healthcare)

Dr. Heather Lynch - Ecology and Evolution Professor, Lynch Lab for Quantitative Ecology

Karen Kernan - URECA Director, Simons Summer Research Program Director

Dr. Eszter Boros - Chemistry Assistant Professor, Boros Lab

Dr. Maria Nagan - Chemistry Lecturer, Nagan Research Lab

Abstract: The recent expansion of online sport wagering and igaming has led to higher rates of problem gambling, particularly among emerging adults and other population subgroups. The Center for Gambling Studies (CGS) at the Rutgers University, School of Social Work, is using big data analysis, machine learning and GIS mapping to identify geographic locations with populations most at risk to guide the development of targeted interventions. This presentation will review the GIS StoryMap for the State of New Jersey, including a blueprint for the highest risk target service areas in the state. It will also present findings from a machine learning model that identifies the key risk factors for high-intensity online casino bettors. Implications for prevention, treatment and policy initiatives will be discussed.

Bio: Lia Nower, J.D., Ph.D., is a Distinguished Professor, Associate Dean for Research, and Director of the Center for Gambling Studies at Rutgers University. A clinician and attorney, her research focuses on big data analysis and machine learning models for online gambling and sports wagering; gambling and video gaming among emerging adults; policy initiatives around harm reduction and responsible gambling, and etiology and treatment of problem gambling. Dr. Nower serves as a senior editor for Addiction. She has received both the Research (2019) and the Lifetime Research Award (2022) from the National Council on Problem Gambling and the Board of Trustees Award for Research (2022) from Rutgers University.

Join Zoom Meeting: https://stonybrook.zoom.us/j/95617197636?pwd=KytzZ2pVRG9SZGpKZUtpNXJISjNjZz09
Meeting ID: 956 1719 7636 Passcode: 924293
Join us at the Center for Excellence in Learning and Teaching (CELT) for an interactive Zoom workshop on Generative AI designed for faculty and staff interested in enhancing teaching and assessment practices, increasing student engagement, and navigating the rapidly evolving landscape of AI tools. Participants will be introduced to common AI tools, explore potential instructional uses, and discuss key considerations such as academic integrity, transparency, and equity.
Register now: https://stonybrook.zoom.us/meeting/register/6js1eP64T1ys8tyU57EJ7Q#/registration