AI + Music Seminar - The meeting will consist of introductions and organizational discussions, aimed at understanding participants' interests. We'll discuss what the seminars can focus on going forward.
As artificial intelligence continues to transform higher education and the world beyond, how are students engaging with this change? Join us for a student-led discussion that explores how AI is influencing academic integrity, learning practices, and students' perspectives on its role in future workplaces.

Our panelists will share their experiences and reflections on questions such as:
1. What counts as appropriate and inappropriate use of AI in coursework?
2. How do faculty approach AI and talk about its implications in class?
3. What does AI mean for students' learning and ethical decision-making?
4. How are students building their understanding of AI tools and their potential uses in professional contexts?

This conversation offers an authentic look at how students are navigating the promises and challenges of AI--both in their studies and as they look ahead to applying these technologies responsibly in their fields.

Register here.

The University at Albany will host a national gathering of professionals and academics that will focus on the transformative potential of AI while addressing the ethical, technical and institutional challenges posed by AI in education.

The Symposium will feature dynamic keynotes, hands-on workshops and engaging conversations with other participants and subject matter experts.

Topics

  • Teaching, Learning and Workforce Development
  • Research, Creative Arts, and Practice
  • Ethics, Governance and Academic Administration

For event information and registration, visit Events@Albany.
https://stonybrook.zoom.us/j/99820812332?pwd=c05BSTVLNmw3L04yZjdEcG5pem1OZz09 Speaker: Alexei Koulakov of Cold Spring Harbor Laboratory Brain evolution as a machine learning problem We have entered a golden age of artificial intelligence research, driven mainly by the advances in ANNs over the last decade or so. Applications of these techniques--to machine vision, speech recognition, autonomous vehicles, machine translation and many other domains--are coming so quickly that many observers predict that the long-elusive goal of Artificial General Intelligence (AGI) is within our grasp. However, we still cannot build a machine capable of building a nest, stalking prey, or loading a dishwasher. I will describe several projects, ranging from theories of evolution of neural development to the perception of smells, in which we are attempting to understand the algorithms that the nervous system is using to solve some of these challenging problems.
CSE 656 Seminar in Computer Vision The overall purpose of this seminar is to bring together people with interests in Computer Vision theory and techniques and to examine current research issues. This course will be appropriate for people who already took a Computer Vision graduate course or already had research experience in Computer Vision. To enroll in this course, you must either: (1) be in the PhD program or (2) receive permission from the instructors. Each seminar will consist of multiple short talks (around 15 minutes) by multiple students. Students can register for 1 credit for CSE656. Registered students must attend and present a minimum of 2 talks. Everyone else is welcome to attend. Fill in https://forms.gle/q6UG9ygauLp2a8Po8 to subscribe to our mailing list for further announcement.
Abstract: Large language models (LLMs) may exhibit unintended or undesirable behaviors. Recent works have concentrated on aligning LLMs to mitigate harmful outputs. Despite these efforts, some anomalies indicate that even a well-conducted alignment process can be easily circumvented, whether intentionally or accidentally. Does alignment fine-tuning yield have robust effects on models, or are its impacts merely superficial? In this work, we make the first exploration of this phenomenon from both theoretical and empirical perspectives. Empirically, we demonstrate the elasticity of post-alignment models, i.e., the tendency to revert to the behavior distribution formed during the pre-training phase upon further fine-tuning. Leveraging compression theory, we formally deduce that fine-tuning disproportionately undermines alignment relative to pre-training, potentially by orders of magnitude. We validate the presence of elasticity through experiments on models of varying types and scales. Specifically, we find that model performance declines rapidly before reverting to the pre-training distribution, after which the rate of decline drops significantly. Furthermore, we further reveal that elasticity positively correlates with the increased model size and the expansion of pre-training data. Our findings underscore the need to address the inherent elasticity of LLMs to mitigate their resistance to alignment.

Speaker: Huajian Zhang

Location: CS2311
Professor Nanpeng Yu from UC Riverside present Machine Learning and Big Data Analytics in Power Distribution Systems.

Abstract: The electric utility industry is being swamped by petabytes of data coming from various sources such as smart meters, phasor measurement units, SCADA systems, geographical information systems and customer management systems. The primary and secondary value embedded in the complex and heterogeneous data sets from power distribution systems is immense. However, algorithms and applications for unlocking the potential of big data in power systems are at an early stage of development. This talk discusses the recent advancement of machine learning algorithms and big data analytics methods in power distribution systems. In particular, we will explain how to develop hybrid algorithms, which synergistically combine the merits of state-of-the-art machine learning algorithms and physical model-based methods. We will take a deep dive into the following applications: network topology identification, electricity theft detection, estimation of behind-the-meter solar generation and data-driven distribution system controls.

Bio: Dr. Nanpeng Yu received his B.S. in Electrical Engineering from Tsinghua University, Beijing, China, in 2006. Dr. Yu received his M.S. degrees in Electrical Engineering and Economics and Ph.D. degree from Iowa State University in 2010. Before joining University of California, Riverside, Dr. Yu was a senior power system planner and project manager at Southern California Edison from Jan, 2011 to July 2014.

Currently, he is an Associate Professor in the Department of Electrical and Computer Engineering at the University of California, Riverside, CA. Dr. Yu is the recipient of the Regents Faculty Fellowship and Regents Faculty Development award from University of California. He received multiple best paper awards from IEEE Power and Energy Society General Meeting, IEEE Power and Energy Society Grand International Conference and Exposition Asia and the Second International Conference on Green Communications, Computing and Technologies.

Dr. Yu is the director of Smart City Innovation Laboratory at UC Riverside. He currently serves as the vice chair of the distribution system operation and planning subcommittee of IEEE Power and Energy Society and the co-chair for IEEE Big Data Applications in Power Distribution Networks Task Force. Dr. Yu currently serves as the associate editor for IEEE Transactions on Smart Grid and International Transactions on Electrical Energy Systems.

The Vedanta Forum is devoted to one of humanity's oldest and most profound pursuits -- thinking. Thinking about who we truly are: the one that remains constant through childhood and old age, through waking, dream, and deep sleep. Thinking about the source and cause of creation, and its relationship to what inheres in us.

Across history, such thinking, both meditative and scientific, has been aimed at these questions. The ancient Upanishads proclaimed, Tat Tvam Asi -- Thou Art That -- revealing the non-dual identity of the individual and the ultimate reality. Centuries later, modern scientists such as Schrödinger and Bohr echoed similar intuitions about the unity of existence.

Over time, many philosophical approaches, traditions, and interpretive schools have arisen from such inquiry, each offering unique perspectives. The Forum will:

  • Focus on universal approaches and traditions and examine their teachings,

  • Foster comparative studies, and

  • Explore the practical benefits to society from such thinking,

through scholarly studies, dialogue, and debate also promoting accessibility to all qualified seekers. Additionally, the Forum will explore how these reflections can enrich life, education, and even technology.

Location: NCS 120 (New Computer Science), Engineering Dr, Stony Brook, NY 11794.

The program is available at: https://www.vedantaforum.org/events/program

Abstract : Humans reason about everyday situations by making commonsense-based inferences, derived both from explicitly stated information and implicit, unstated knowledge. In this thesis, I investigate whether NLP models have different aspects of causal knowledge about events and how to improve their understanding of narratives and plans.
Answering questions about why people perform actions in a narrative can test whether NLP systems contain and can effectively apply causal knowledge about events. I introduce TellMeWhy, a dataset concerning why characters in short narratives perform the actions described. An evaluation of then SOTA finetuned models show that they are far worse than humans. To improve models, it is important to understand what aspects of causal knowledge they need and how to best use external sources to inject this knowledge. In KnowWhy, I analyze different ways of injecting knowledge into models, which is difficult since we do not know apriori what type of knowledge will be needed to answer a question, hence requiring a ranking model to pick the most important inference. Results show that this retrieved knowledge helps models of all sizes, thereby improving their understanding of narratives.
Next, I study whether models can reason about causal aspects of plans. I focus on testing whether they understand the underlying causal dependencies reflected in the temporal order of a plan's steps. I introduce CAT-Bench, and find that SOTA models are underwhelming, and that model answers are not consistent across questions about the same step pairs. In their current state, these models cannot yet reliably be used for complex user-facing tasks. I then measure contemporary models' ability to perform user-facing and user-centric plan customization. I introduce the use of semi-symbolic edits in large language model (LLM) based agents and test several multi-LLM-agent architectures for plan customization. While LLMs still lack the ability to understand complex customization hints, my results suggest that LLM-based architectures may be worth exploring further for other customization applications. Finally, I distill complex reasoning capabilities into small language models (SLMs) using synthetic data that reflects a decomposition-then-editing process for plan customization. I demonstrate that explicitly teaching this latent causal reasoning significantly improves the quality of SLM-generated customizations. Overall, my work has improved how well NLP models understand complex reasoning associated with events in different contexts.

Speaker: Yash Kumar Lal

Location: NCS 220 or Zoom https://stonybrook.zoom.us/j/95849648243?pwd=dgPpZtDpgwQrK9z1SaPpNbBifaorzk.1