Description:

As artificial intelligence and data science reshape the global information landscape, libraries are emerging as key players in both technological innovation and ethical stewardship. This international Zoom discussion brings together library professionals and educators from the U.S., Philippines, and Hong Kong to explore how institutions are integrating AI and data into their pedagogy and services.

Panelists will share concrete examples from their own libraries--ranging from data literacy initiatives to increasing discoverability. The conversation will also examine regional trends in librarianship, spotlighting how institutions in Asia are navigating the evolving role of data and AI.

Join us for a global conversation that highlights the transformative potential of libraries as hubs for innovation and critical inquiry in the age of AI.

Register for this free Zoom panel.

Panelists:

Ahmad Pratama is a Faculty Member and Associate Librarian at Stony Brook University Libraries, where he is working to build a comprehensive, campus-wide data literacy program within the Libraries. As the Data Literacies Lead, his work focuses on empowering students, faculty, and staff to critically and ethically engage with data and AI, including the development of a credit-bearing course in Critical Data & AI Literacies supported by an EDGE Fund Award from the Provost's Office. Previously, Dr. Pratama served as an Associate Professor of Information Technology, and his research and teaching explore the intersections of technology, policy, and society with a focus on data, AI, and innovation in higher education.

Dan Anthony Dorado is a full-time faculty member at the U.P. School of Library and Information Studies, where he teaches information technology, management and marketing, research methodology, and quantitative research. He was also the director of the Diliman Learning Resource Center under the Office of the Vice Chancellor for Student Affairs. Before that, he was an Information Specialist at the College of Engineering Library, in charge of the System and Network Administration and The Learning Commons. He completed his master's degree at the Technology Management Center in U.P. Diliman and is currently pursuing his PhD in Data Science. As a member of Sync.Bio.Optics laboratory and the Publics, Archives, and Data (PANDA) Lab, his research specialization covers Computational Methods, Open Education, Critical Data Studies, and Radical Statistics.

Ryun LEE is Associate University Librarian at The Chinese University of Hong Kong Library, leading Digital Initiatives and Library IT and Systems. He drives digital innovation through emerging technologies, particularly artificial intelligence to enhance services, streamline operations, and support CUHK's mission in research, education, and knowledge advancement. With a background in cataloging and digital repository development, Ryun leads projects in digitization, OCR, data visualization, text and network analysis, GIS, and digital scholarship. He actively promotes knowledge graph applications in Hong Kong studies and oversees efforts to digitize and preserve resources related to Hong Kong and Southern China. His recent work focuses on creating seamless digital experiences and developing data-driven infrastructure. He is currently exploring AI-driven approaches to digitization workflows and entity extraction, aiming to improve access, discovery, and long-term preservation of library materials.











Abstract:
Quantifying similarity is a central notion in science and data analysis, pervading everything from phylogenetic trees to the foundation of clustering. Unfortunately, despite being examined and applied for decades, traditional similarity and distance metrics have fundamental drawbacks. The key problem is that all of them are only defined over pairs of objects, so they scale quadratically when one tries to compare N objects. The present explosion in the amount of data available to us requires new ways to process information, and while some current algorithms can handle millions of points, we need alternatives applicable to billions. This is what motivated us to develop a new framework that can compare any number of objects at the same time. With this, we achieve an unprecedented linear scaling when comparing multiple objects. Here we will discuss the main properties of this formalism, along with its applications in drug design and to the analysis of Molecular Dynamics (MD) simulations. Our indices have proven to be incredibly versatile when applied to chemical space exploration and visualization, allowing us to rigorously quantify the chemical diversity of very large molecular libraries. This has led to the creation of several algorithms to sample important regions in chemical space, including a more efficient way of identifying the prevalence of activity cliffs. Additionally, our indices provide a convenient route to sample complex MD trajectories, allowing to identify representative structures very efficiently. Moreover, we can also cluster biological ensembles in a more robust way than with standard algorithms, which has led to our group's work on MDANCE, a very flexible and efficient open-source clustering module. Drop by if you want to know how we clustered one billion molecules!


Speaker:
Assistant Professor, Department of Chemistry and Quantum Theory Project
University of Florida, Gainesville
Website: https://quintana.chem.ufl.edu/

Location:
Laufer Center Lecture Hall 101
CSE 600 Talk: Squeezing Software Performance via Eliminating Wasteful Operations presented by Xu Liu

ABSTRACT: Inefficiencies abound in complex, layered software. A variety of inefficiencies show up as wasteful memory operations, such as redundant or useless memory loads and stores. Aliasing, limited optimization scopes, and insensitivity to input and execution contexts act as severe deterrents to static program analysis. Microscopic observation of whole executions at instruction- and operand-level granularity breaks down abstractions and helps recognize redundancies that masquerade in complex programs. In this talk, I will describe various wasteful memory operations, which pervasively exist in modern
software packages and expose great potential for optimization. I will discuss the design of a fine-grained instrumentation-based profiling framework that identifies wasteful operations in their contexts, which guides nontrivial performance improvement. Furthermore, I will show our recent improvement to the profiling framework by abandoning
instrumentation, which reduces the runtime overhead from 10x to 3% on average. I will show how our approach works for native binaries and various managed languages such as Java, yielding new performance insights for optimization.

BIO: Xu Liu is an assistant professor in the Department of Computer Science at College of William & Mary. He obtained his PhD from Rice University in 2014 and joined the College of William & Mary in the same year. Prof. Liu works on building performance tools to pinpoint and optimize inefficiencies in HPC code bases. He has developed several open-source profiling tools, which are used worldwide at universities, DOE national laboratories and industrial companies. Prof. Liu has published a number of papers in high-quality venues. His papers received Best Paper Award at SC'15, PPoPP'18, PPoPP'19 and ASPLOS'17 Highlights, as well as Distinguished Paper Award at ICSE'19. His recent ASPLOS'18 paper has been selected as ACM SIGPLAN Research Highlights in 2019 and nominated for CACM Research Highlights. Prof. Liu is the receipt of 2019 IEEE TCHPC Early Career Researchers Award for Excellence in High Performance Computing. Prof. Liu served on the program committee of conferences such as SC, PPoPP, IPDPS, CGO, HPCA and ASPLOS.

This course is designed to help you approach AI with clarity and intention. Explore how to ask better questions, apply critical judgment to AI-generated outputs, and use these tools to support--not replace--human decision-making. With the right prompts, a thoughtful lens, and a focus on impact, AI can help reduce friction and free you to focus on what matters most: people, purpose, and leadership.

Audience: Any current or aspiring leader who currently uses AI and is interested in bringing their skills to the next level.

Register here.

Abstract: This dissertation addresses the methodological disconnect between Natural Language Processing (NLP) and human-centric analysis by shifting the unit of analysis from document to human behavior in two broad respects: (i) time-ordering: modeling documents as sequential person-indexed behavioral observations, and (ii) person-level semantics: evaluation and explainability of models by their latent structure of psychological constructs rather than just its predictive accuracy against narrow proxy measures. First, we consider the most basic implication of language as a person's behaviors when measuring their psychological constructs: relationship between language sample size and model's predictive performance. We empirically show that the state-of-the-art transformers are often over-parameterized for typical NLP dataset sizes and can be reduced in dimensionality without performance loss. Establishing the author as the unit of analysis naturally allows us to treat their behavior as a time-ordered sequence. Second, we introduce a longitudinal evaluation framework that establishes ecologically valid evaluation settings, namely, cross-sectional and prospective generalization, and separates error measurement of the model into within-person dynamics and between-person differences. We demonstrate that traditional NLP evaluations based on random document splits can yield reversed conclusions under ecologically valid generalization settings. To address this, we develop models that capture the trajectory of mental states (e.g., mood shifts) rather than static traits. Third, moving into person-level semantics, we evaluate the latent structure of large language models using a novel machine behavior analytic framework. We find that while GPT-4 achieves high predictive correlation with self-reports, its latent symptoms structure diverges from clinical understanding. Finally, we propose a method for modeling multidimensional behaviors, embedding concurrent behavioral signals alongside language to predict future states. Taken together, this work suggests that operationalizing language as behavior advances NLP methods into a rigorous instrument for valid psychological inquiry.

Speaker: Adithya Ganesan

Location: Join Zoom Meeting (ID: 99021939129, Passcode: 569493)

You are cordially invited to attend the biweekly Brookhaven AI Mixer (BAM). BAM includes one short talk on AI research happening at BNL, followed by an open mixer over coffee and snacks for everyone to network and discuss all things AI. The first half hour will consist of presentations that will be available via ZOOM, and the second half hour will be for in person only networking.

Join us every other Tuesday at noon in CDSD's Training Room (building 725, 2nd floor) to learn about interesting AI methods and applications, engage with potential collaborators, prepare for pending FASST funding calls, and build a community of AI for Science at BNL.

#1 How to train your Scientific Chatbot by Alexandr Prozorov, Post-Doctoral Research Associate


Abstract: RHIC is closing its 25-year run with ~1 EB of data and decades of hard-won know-how that risk drifting into obscurity. The RHIC Data & Analysis Preservation Plan (DAPP) pilots an AI assistant that lets physicists talk to RHIC in natural language--searching internal notes, code, workflows, and docs, and pointing to runnable, containerized analyses. Built on Retrieval-Augmented Generation(RAG) with a Model Context Protocol orchestration layer, the system indexes heterogeneous, experiment-specific content and enforces role-aware access
for public vs. collaboration-restricted materials. Takeaway: domain-adapted AI can turn a legacy exabyte into reproducible answers, training assets, and new discovery paths.

Biography: Alexandr Prozorov is a postdoc from Czech Technical University in Prague working in STAR experiment. Fascinated by AI

#2 Quantum AI: Atoms, Cavities and Learning by Raman Kumar, Post-Doctoral Research Associate, Instrumentation Department

Abstract: The Instrumentation Department (IO) in the Discovery Technologies directorate at BNL is engaged in exploring various aspects of quantum systems research. One of the main goals of our group's effort is in developing neutral atom-cavity array platforms for remote entanglement generation and distributed quantum processing. This platform promises to herald truly scalable quantum computing systems and open new paradigms for networking and sensing. In this talk, I will explain our group's research and the role AI is playing in unlocking new insights with two examples. The first application of AI is in fabrication process prediction of micro-cavity structures. The second application revolves around role of AI in quantum error detection and correction in modern quantum computing systems.

Biography: Dr. Raman Kumar is a postdoctoral research associate in the IO department at BNL working with Dr./Prof. Sebastian Will (Columbia U.). Kumar obtained his Ph.D. degree in Electrical and Computer Engineering from the University of Illinois Urbana-Champaign. Prior to joining BNL in Nov 2024, Kumar worked as a postdoc at the City College in New York working on topological photonic quantum sensing using NV centers in diamond. Kumar and Will combined have an extremely wide moat and expertise in a variety of different areas which include Ultra cold atoms and molecules, quantum optics, quantum condensed matter, nanofabrication, semiconductor devices and advanced electromagnetics. Their areas of research interest include scalable quantum computing, communications and sensing, all enabled by AI.

Location: CDS, Bldg. 725, Training Room

Join ZoomGov Meeting https://bnl.zoomgov.com/j/1607892208?pwd=MSjxN5btSeToZsQMwEQzCCbBo5h58V.1

Meeting ID: 160 789 2208
Passcode: 753871

University Libraries Presents: The Library AI Club is a welcoming space for students, faculty, and staff to explore AI in a supportive, low-pressure environment. Meeting every two weeks, the club features discussions, collaborative projects, guest speakers, and hands-on experiments. Join us to learn, share ideas, and engage with AI responsibly and creatively. We'd love to see you at an upcoming meeting! Location: Melville Library, Scholarly Communication Seminar Room