2025 AGENDA
Speaker/Presenting Author in Italics
Monday, September 15
1-K: Keynote Session (10:30-11:00)
Co-Chairs: J. Kepner & A. Reuther
- Keynote Talk: Enabling Advances in OceanAI
- Nick Rotker (MITRE)
1-1: Advanced Processor Architectures Session (11:00-12:15)
Co-Chairs: P. Luszczek & C. Byun
- Invited Talk:
- Michael Foertsch (Q.ANT)
- Encoded Time-Series Model Training with UNet Running on Wafer Scale Engine
- Vyacheslav Romanov (NETL)
- Lincoln AI Computing Survey (LAICS) and Trends
- Albert Reuther, Peter Michaleas, Michael Jones, Vijay Gadepally, Jeremy Kepner (MIT Lincoln Laboratory Supercomputing Center)
- System-Level Performance Modeling of Photonic In-Memory Computing
- Jebacyril Arockiaraj, Sasindu Wijeratne (USC), Sugeet Sunder (USC Information Sciences Institute), Md Abdullah-Al Kaiser, Akhilesh Jaiswal (Univ. of Wisconsin), Ajey P Jacob (USC Information Sciences Institute), Viktor Prasanna (USC)
1-P1 (12:15-13:15): New Application Frontiers Poster Session
Chair(s)/Host(s): K. Keville
- Multi-Stage Stochastic Programming for Heavy-Duty Electric Truck Routing Under Public Charging Congestion Uncertainty
- Ziyan Li, Nikolay Aristov, Antoine Germain, Elenna R. Dugundji (MIT)
Enabling Heterogeneous Performance Analysis for Scientific Workloads [Outstanding Short Paper Award]
- Maksymilian Graczyk (CERN), Vincent Desbiolles (HES-SO), Stefan Roiser (CERN), Andrea Guerrieri (HES-SO and EPFL)
- Weighted Histogram Matching for Improved Automated Synthetic Aperture Radar (SAR)-optical Image Registration
- Kerri Prinos, Aditi Mungale, Dr. Karen Gettings, Dr. Tair Akhmejanov, Devanshu Mehta (MIT Lincoln Laboratory)
- Towards a Framework for Etymological Analysis of Literary and AI Works
- Mike Teodorescu (Univ. of Washington), Cecilia Speranța Bolea (IIT), Horia Teodorescu (Technical Univ. Iasi)
An Algorithmically Determined Exchange-Correlation (ADEXC) Method for Electronic Structure Calculations
- Ivan Williams, Eric Polizzi (UMass Amherst)
Scaling Performance of Large Language Model Pretraining [Best Short Paper Award]
- Alexander Interrante-Grant, Carla Varela-Rosa, Suhaas Narayan, Chris Connelly, Albert Reuther (MIT Lincoln Laboratory)
Metadata Guided Pose Estimation for 3D Reconstruction
- William Farthing (Emory Univ.), Compton Ross (Univ. of Mississippi), Jaired Collins, Jing-Ru C. Cheng (U.S. Army Corp of Engineers)
Neuromorphic Processor Employing FPGA Technology with universal interconnections
- Pracheta Harlikar, Abdel-Hameed Badawy (New Mexico State Univ.), Prasanna Date (ORNL)
1-2: Neuromorphic AI Session (12:30-13:45)
Co-Chairs: M. Barnell & H. Nguyen
- Artificial Intelligence Performance and Radiation Effects in Neuromorphic NorthPole Hardware
- Victor M. Vergara (BlueHalo-AFRL), Francisco O. Viramontes, Amanda E. Romero (COSMIAC Research Center), Matthew E. Spear, Evan T. Kain, Windy S. Slater, Heather M. Quinn, Qing Wu (AFRL)
- Physics-Informed Neural Networks for Low-Power Real-Time Edge Biosensing Applications
- Soheli Farhana (Harvard Univ.)
- Analysis and Optimization of Spiking Neural Network Simulations on GPUs
- Doğu Kocatepe, Işıl Öz (Izmir Institute of Technology)
- Exploring Neuromorphic Computing with Loihi-2 for High-Performance CFD Simulations
- Talha Coskun (Univ. of Illinois Urbana-Champaign), Hiruna Vishwamith (Univ. of Moratuwa), Murat Isik (Stanford Univ.), I. Can Dikmen (Istinye Univ.)
- Neural-Inspired Enhancing Spiking Graph Convolutional Networks
- Fernando Vera Buschmann, Horacio Roststein, Vincent Oria (New Jersey Inst. of Tech.)
1-P2 (13:45-14:45): High Performance Data Analysis Poster Session
Chair(s)/Host(s): K. Keville
- Accelerating Push-Relabel Algorithm on GPU via Two-Level Parallelism Paradigm and Efficient CSR Designs
- Chou Ying Hsieh, Po-Chieh Lin, Sy-Yen Kuo (National Taiwan Univ.)
- Improving Statistical Characterization of Data Tensors with the Generalized Canonical Polyadic Tensor Decomposition
- Matthew Merris, Tim Andersen (Boise State Univ.)
Comparative Analysis of Classical and Deep Learning Features for Texture Image Classification
- Murad Hossen (Univ. of Houston)
Aspects of HPC within Conformal Recommendation Systems: Ideas and Challenges
- Stanisław M. S. Halkiewicz (AGH Univ. of Cracow), Maciej Kuczyński (Czestochowa Univ. of Tech.)
- Enhancing the Real-Time Solutions of Parametric Linear Systems on a GPU through Hybrid Coarse-Grained Transprecision Computing
- Hamid Noori, Hans Vandierendonck, Roger Woods (Queen’s Univ. Belfast), Nick Polydorides (Univ. of Edinburgh)
AAPA: An Archetype-Aware Predictive Autoscaler with Uncertainty Quantification for Serverless Workloads on Kubernetes
- Guilin Zhang (George Washington Univ.), Srinivas Vippagunta, Raghavendra Nandagopal, Suchitra Raman, Jeff Xu, Marcus Pfeiffer, Shreeshankar Chatterjee (Workday), Ziqi Tan, Wulan Guo (George Washington Univ.), Hailong Jiang (Youngstown State Univ.)
- Stable Iterative Solvers for Ill-Conditioned Linear Systems and Least Squares
- Vasileios Kalantzis, Mark S. Squillante, Chai Wah Wu (IBM Research)
1-3: ASIC and FPGA Advances Session (14:15-15:30)
Co-Chairs: L. Zaidenberg & K. Thurmer
- NTT-SAA: Exploring NTT Acceleration with 2-D Systolic Array Architecture on FPGAs
- Ashwajit Singh (IIT Bombay), Zhihan Xu, Viktor K. Prasanna (USC)
- Evaluating AMD-Xilinx Frameworks for Deep-Learning Acceleration on Versal [Outstanding Student Paper Award]
- Peter Drum, Alan D. George (NSF SHREC)
- Fast FPGA-Based Implementation of the QP-Dyn Stream Cipher using High-Level Synthesis
- Paolo Palazzari (ENEA), Luigi Accardi (Volterra Univ.), Antonio Mastrandrea, Pasquale Tommasino, Alessandro Trifiletti (Sapienza Univ.)
- A Scalable Code Generation Flow for Heterogeneous Parallel RTL Simulation using MLIR
- Jie Tong, Zhengxiong Li, Umit Yusuf Ogras, Tsung-Wei Huang (Univ. of Wisconsin)
- Accelerating Dynamic Image Graph Construction on FPGA for Vision GNNs
- Anvitha Ramachandran, Dhruv Parikh, Viktor Prasanna (USC)
1-4: FastCode Session (15:45-17:15)
Co-Chairs: Bruce Hoppe
- Invited Talk: Fastcode: An Open-Source Community for Making Software Performance Engineering Easy and Fun
- Bruce Hoppe (MIT)
- Invited Talk: Taskflow — A General-Purpose Task-Parallel Programming System
- Tsung-Wei Huang (Univ. of Wisconsin)
- Invited Talk: OpenCilk — A Modular and Extensible Software Infrastructure for Fast Task-Parallel Code
- I-Ting Angelina Lee (WUSTL)
- Invited Talk: Multithreaded Parallel Python Through OpenMP Support in Numba
- Tim Mattson (Human Learning Group)
1-5: BRAINS – Building Resilience through Artificial Intelligence for Networked Systems Session (17:30-19:30)
Organizers: T. Hardjono & S. Pisharody
- Invited Talk: Developing Trust in the Supply Chain
- Guy Fedorkow (Juniper/HPE)
- Invited Talk: Using AI to Achieve Peak Supply Chain Health
- Karyl Fowler (Tradeverifyd)
- Invited Talk: Supply Chain Transparency as Entailment for Device Attestation
- Ned Smith (Intel)
- Towards an Algorithm-based Approach for Soft Error Tolerance using Interval Arithmetic [Best Paper Award]
- Larry Tang, Varun Kumar, Matt Ngaw, Siddharth Singh, Devdutt Nadkarni, Lohith Tummala, Ken Mai, Franz Franchetti (Carnegie Mellon Univ.)
- Detecting Expert-Written Comments in Stack Exchange
- Himani Musku (Carnegie Mellon Univ.), Alea Ritchie (Stanford Univ.), Nour Jedidi, Rohan Leekha, Courtland VanDam (MIT Lincoln Laboratory)
- Strategic Cyber Defense via RL-Guided Combinatorial Auctions
- Mai Pham, Vikrant Vaze, Peter Chin (Dartmouth Coll.)
- AI-Powered Network Energy Optimization Machine Learning Approaches to Reducing Network Power Consumption
- Het Mehta (Cisco Systems), Sambu Patach Arrojula (Samsung)
Tuesday, September 16
2-K: Keynote Session (10:30-11:00)
Co-Chairs: J. Kepner & A. Reuther
- Keynote Talk: Effective Human-Machine Partnerships in High Stakes Settings
- Julie Shah (MIT)
2-1: Bridging Quantum and HPC Session (11:00-12:15)
Co-Chairs: Devesh Tiwari
- Invited Talk: Tightening the Integration of Quantum Computers with HPC Systems
- Travis Humble (ORNL)
- Invited Talk: Enabling Quantum Utility through a System Toolkit
- Tirthak Patel (Rice Univ.)
- Invited Talk: CUDA-Q: Defining a Tightly-Integrated System Software Stack for Quantum-Accelerated HPC
- Alex McCaskey (NVIDIA)
2-P1 (12:15-13:15): Graph AI & Sparse Poster Session
Chair(s)/Host(s): K. Cain
- Cyber Orbits of Large Scale Network Traffic
- Jeremy Kepner, Hayden Jananthan, Chasen Milner, Michael Houle, Michael Jones, Peter Michaleas, Alex Pentland (MIT)
- Accelerating AI Development with Cyber Arenas
- William Cashman, Chasen Milner (USAF), Michael Houle, Michael Jones, Hayden Jananthan, Jeremy Kepner, Peter Michaleas, Alex Pentland (MIT)
- Optimizing Sparse Matrix-Vector Multiplication on GPUs using the Mathematics of Arrays
- Stephen Thomas (Lehigh Univ.), Lenore Mullin (Univ. of Albany)
Sampling to Scale: Performance Trade-offs in Approximate Triangle and Square Counting
- Shubhashish Kar, Shaikh Arifuzzaman (UNLV)
Spectral Sparsification of Edges for Efficient Clique Counting
- Aaron Schindler (Univ. of Utah), Hari Sundar (Tufts Univ.)
- Degree Matrix Comparison for Graph Alignment
- Ashley Wang, Peter Chin (Dartmouth Coll.)
2-2: Case Studies, Benchmarking, and Tools Session (12:30-13:45)
Co-Chairs: X. Sun & J. Ghanem
- Invited Talk: Rigorous Mathematics in Data-driven Methods for Safe Autonomy
- Chuchu Fan (MIT)
- AOCL-CST – A new CPU Stress Test library for AMD CPUs
- S. Biplab Raut (AMD)
- Performance–Energy Characterization of ML Inference on Heterogeneous Edge AI Platforms
- Palash Kohli (BITS Pilani), Rakshith Jayanth, Neelesh Gupta, Haoyang Fan, Viktor Prasanna (USC)
- A Fully Adaptive Radau Method for the Efficient Solution of Stiff Ordinary Differential Equations at Low Tolerances
- Shreyas Ekanathan (Lexington High School), Oscar Smith (JuliaHub), Christopher Rackauckas (MIT)
- Evaluation of Habitat Robotics using LargeLanguage Models
- William Li, Lei Hamilton, Kaise Al-natour, Sanjeev Mohindra (MIT Lincoln Laboratory)
2-3: Quantum and Non-Deterministic Computing Session (14:15-15:30)
Co-Chairs: B. Sroka & I. DeTore
- SYCL for Performance Portability: Application Experience with Coupled Cluster Formalism in Quantum Chemistry on Exascale Systems [Outstanding Paper Award]
- Abhishek Bagusetty (Argonne National Laboratory), Ajay Panyala (PNNL), Alvaro Vazquez Mayagoitia (Argonne National Laboratory), John K. Holmen (ORNL), Kevin Harms (Argonne National Laboratory)
- Implementation of Tensor Network Simulation TN-Sim under NWQ-Sim
- Aaron C. Hoyt, Jonathan S. Bersson, Sean Garner (Univ. of Washington), Chenxu Liu, Ang Li (PNNL)
- Partition-based Surface Code Compilation
- Hanjing Xu (Purdue Univ.), Xiaoyuan Liu, Ankit Kulshrestha, Hayato Ushijima-Mwesigwa (Fujitsu Research of America)
- QAOA Parameter Transferability for Maximum Independent Set using Graph Attention Networks
- Hanjing Xu (Purdue Univ.), Xiaoyuan Liu (Fujitsu Research of America), Alex Pothen (Purdue Univ.), Ilya Safro (Univ. of Delaware)
- A Hybrid Classical-Quantum Model for QSAR-Based Biodegradability Prediction
- Batuhan Hangun, Oguz Altun (Yıldız Technical Univ.), Onder Eyecioglu (Bolu Abant İzzet Baysal Univ.)
2-4: Graph AI & Sparse Session (15:45-17:15)
Co-Chairs: X. Sun & J. Ghanem
- Scalable Graph Algorithms on Distributed UpDown Accelerators
- Brian Wheatman, Andrew A. Chien (Univ. of Chicago)
- 2D Distributed Label Propagation on 400 GPUs
- George M. Slota, Michael Mandulak, Ujwal Pandey, Anthony Fabius (RPI)
- Accelerating Sparse Deep Learning via Multi-Layer Tensor Reordering and Partitioning
- Gunduz Vehbi Demirci (Imagination Tech.), Cagatay Dikici (Wayve), Tim Atherton (Imagination Tech.)
- Julia GraphBLAS with Nonblocking Execution [Outstanding Paper Award]
- Pascal Costanza (Independent Researcher), Timothy G. Mattson (Univ. of Bristol), Raye Kimmerer (NERSC LBNL), Benjamin Brock (Intel)
- A parallel push-relabel maximum flow algorithm in LAGraph and GraphBLAS
- Darin A. Peries, Timothy A. Davis (Texas A&M Univ.)
- GraphBLAS Mathematical Opportunities: Parallel Hypersparse, Matrix Based Graph Streaming, and Complex-Index Matrices
- Hayden Jananthan, Jeremy Kepner, Michael Jones, Vijay Gadepally, Michael Houle, Peter Michaleas, Chasen Milner, Alex Pentland (MIT)
2-5: GraphBLAS BoF Session (17:30-19:30)
Organizers: T. Mattson, B. Brock & S. McMillan
- GraphBLAS APIs: The Math Spec
- Tim Mattson (Human Learning Group)
- Keynote – Postgres and GraphBLAS
- Michel Pelletier (OneSparse)
- SuiteSparse GraphBLAS and LAGraph: Progress Report and Future Direction
- Tim Davis (Texas A&M)
Wednesday, September 17
3-K: Keynote Session (10:30-11:00)
Co-Chairs: J. Kepner & A. Reuther
- Keynote Talk: Building AI That Users Trust
- Ashley Conard (Microsoft)
3-1: Scaling Research Computing Education Session (11:00-12:15)
Co-Chairs: J. Mullen, L. Milechin & H. Jananthan
- Invited Talk: Research/Advanced Computing Roles: Skillfully Facing the Challenges
- Robert M. Freeman, Jr. (Harvard University & CaRCC)
- Invited Talk: Skill Inventories, What They Are and Why We Need Them
- Weronika Filinger (Edinburgh Parallel Computing Center)
- Invited Talk: Cataloguing the Training Material Landscape: Skills, Gaps, and Resources
- Jeremy Cohen (Imperial Col.)
- Invited Talk: Designing an Integrated Ecosystem to Develop Research Computing and Digital Technical Professionals
- Julia Mullen (MIT LL)
3-2: AI in the World Session (12:30-13:45)
Co-Chairs: K. Gettings & M. Barnell
- Invited Talk: Learning-Guided Optimization for Mobility
- Cathy Wu (MIT)
- Sustainably Modeling a Sustainable Future Climate
- Rabab Alomairy (MIT), Sameh Abdulah (KAUST), Qinglei Cao (St. Louis Univ.), Marc G. Genton, David E. Keyes, Hatem Ltaief (KAUST)
- Predicting Ports Congestion by Utilizing State Space Models
- Nikolay Aristov, Elenna R. Dugundji (MIT-CTL)
- Scaling Regime-Aware Forecasting: Distributed Shifting Seasonal Matrix Factorization
- Jacob Munson, Breschine Cummins (Montana State Univ.)
- AIMS: An Adaptive Intelligent Multi-Objective Scheduler Powered by Digital Twins
- Kyrian Adimora, Hongyang Sun (Univ. of Kansas)
3-P2 (13:45-14:45): High Performance Computing Poster Session
Chair(s)/Host(s): P. Luszczek
- When Structure is Silent: Opportunities for Algorithmic Dispatch in Linear Algebra
- Emmanuel Lujan, Alan Edelman (MIT)
- Performance Modeling of Heterogeneous Edge-Cloud Systems with Machine Learning
- Md Raihan Uddin, Abu Asaduzzaman, Sonu Gangadhar Gowda (Wichita State Univ.)
- An FFT-based Preconditioner for Conjugate Gradient Pressure Solvers in Complex Domains
- Xin Kai Lee, Gregory LeClaire Wagner, Simone Silvestri, Raffaele Ferrari (MIT)
- A Scalable Quantum Dynamical Approach for Calculating Collisional Molecular Properties
- Prajwal Niraula, Laurent Wiesenfeld, Julien de Wit (MIT), Iouli Gordon, Robert Hargreaves (Harvard Univ.), Jeremy Kepner, Deborah Woods, Cooper Loughlin (MIT Lincoln Laboratory)
- Scalable Bayesian Nonparametric Ensemble (BNE) for Spatio-temporal Air Pollution Predictions using High-Performance Computing
- Vijay Kumar, Jaime Benavides (Brown University), Carlos Carrillo-Gallegos (Columbia Univ.), Gil Speyer (Arizona State Univ.), Marianthi-Anna Kioumourtzoglou (Brown University)
- Investigating the Impact of Algorithms and Hardware on Machine Learning Models in HPC Systems
- Christian C. Thompson, Abu Asaduzzaman, Md Raihan Uddin (Wichita State Univ.)
- SARComp: High-Performance Algorithms for Onboard SAR from FFT Kernels to Matched Filtering
- Maron Schlemon (German Aerospace Center), Martin Schulz (Tech. Univ. Munich), Rolf Scheiber (German Aerospace Center)
Streamed Multi-Format Sparse Matrix Vector Multiplication for FPGA
- Spencer Smith, Richard Veras (Univ. of Oklahoma)
3-3: Graph AI & Sparse Session (14:15-15:30)
Co-Chairs: N. Pitsianis & C. Byun
- pdGRASS: A Fast Parallel Density-Aware Algorithm for Graph Spectral Sparsification [Best Student Paper Award]
- Tiancheng Zhao (Georgia Inst. of Tech.), Zekun Yin, Huihai An (Shandong Univ.), Xiaoyu Yang (China Univ. of Petroleum-Beijing), Zhou Jin (Zhejiang Univ.), Jiasi Shen (HKUST), Helen Xu (Georgia Inst. of Tech.)
- Differentiable Graph Centrality
- Georgios Kollias, Vassilis Kalantzis (IBM Research)
- Performance Analysis of the Parallel Shared-Memory Sparse Matrix-Vector Multiplication on Unstructured Matrices
- Kobe Bergmans, Karl Meerbergen, Raf Vandebril (KU Leuven)
- HiPerMotif: Novel Parallel Subgraph Isomorphism in Large-Scale Property Graphs
- Mohammad Dindoost, Oliver Alvarado Rodriguez, Bartosz Bryg, Ioannis Koutis, David A. Bader (New Jersey Inst. of Tech.)
- Characterization of Sparsity-aware Parallelization of Jaccard Similarity in Graph Datasets
- Atharva Gondhalekar, Paul Sathre, Wu-chun Feng (Virginia Tech)
3-4: Graph AI and Sparse Session (15:45-17:15)
Co-Chairs: N. Pitsianis & C. Byun
- Enhancing Graph Partitioning with Reinforcement Learning-based Initialization
- Chedi Morchdi (Texas A&M Univ.), Cheng-Hsiang Chiu, Wan Luan Lee, Tsung-Wei Huang (Univ. of Wisconsin), Yi Zhou (Texas A&M Univ.)
- Benchmarking Deep Learning with Representative ONNX Subgraphs
- Marika E. Schubert (Univ. of Pittsburgh), David Langerman (NSF SHREC), Evan W. Gretok, Ian Peitzsch, Calvin B. Gealy, Jefferson Boothe, Alan D. George (Univ. of Pittsburgh)
- A Locality Sensitive Hashing Based Algorithm to Accelerate Neighborhood Search in Graph Neural Operators
- Mariam Hassan, Sanmukh Kuppannagari (Case Western Reserve Univ.)
- Power Iteration with Probabilistic Updates for Systems with Heterogeneous Performance
- Soumyadip Ghosh, Lior Horesh, Vasileios Kalantzis, Georgios Kollias, Yingdong Lu, Tomasz Nowicki, Shashanka Ubaru (IBM Research)
- Designing Parallel Algorithms for Community Detection using Arachne
- Fuhuan Li, Zhihui Du, David Bader (New Jersey Inst. of Tech.)
- GCN-Driven CUDA Parameter Optimization for Parallel Triangle Counting in Graphs
- Hasan Serdar Arikan, Rakibul Hassan, Shubhashish Kar (Univ. of Nevada Las Vegas), Doru Thom Popovici (LBNL), Shaikh Arifuzzaman (Univ. of Nevada Las Vegas)
3-5: Graph Challenge Session (17:30-19:30)
Organizers: J. Kepner & A. Reuther
- Combining Performance and Productivity: Accelerating the Network Sensing Graph Challenge with GPUs and Commodity Data Science Software
- Siddharth Samsi, Dan Campbell, Emanuel Scoullos, Oded Green (NVIDIA)
- SANST: Sensing Anonymized Network via Sorted Triplets
- Jianyu Wang, Wenzi Tang, Chenglong Shi, Zhe Zhang (Guangxi Univ.), Dan Chen (National Univ. Singapore), Miaojiang Chen, Wenjing Xiao (Guangxi Univ.)
- Anonymized Network Sensing using C++26 std::execution on GPUs
- Michael Mandulak (RPI), Sayan Ghosh, S M Ferdous, Mahantesh Halappanavar (PNNL), George Slota (RPI)
- Geans: A GPU-accelerated Framework for Efficient End-to-End Anonymized Network Sensing
- Jun Mai, Qinggang Wang, Yu Huang, Pengcheng Yao, Long Zheng, Xiaofei Liao, Hai Jin (Huazhong Univ. of Science and Tech.)
- DBOS Network Sensing: A Web Services Approach to Collaborative Awareness
- Sophia Lockton, Jeremy Kepner, Michael Stonebraker, Hayden Jananthan, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel Burrill, Chansup Byun (MIT), Timothy Davis (Texas A&M), Vijay Gadepally, Michael Houle, Matthew Hubbell, Michael Jones, Piotr Luszczek, Peter Michaleas, Lauren Milechin, Chasen Milner, Guillermo Morales, Julie Mullen (MIT), Michel Pelletier (OneSparse), Alex Poliakov (DBOS), Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Alex Pentland (MIT)
- Interactive Trillion Packet Anonymized Network Analysis with the GraphBLAS
- Chasen Milner (USAF), Michael Houle, Hayden Jananthan, Michael Jones, Jeremy Kepner, Peter Michaleas, Inna Voloshchuk, Alex Pentland (MIT)
- Towards Efficient Sparse Deep Neural Network Inference via Multi-level Concurrency Orchestration
- Ming Dun, Jie Zhou (Inst. of Computing Tech, CAS), Huawei Cao (UCAS), Shuhan Song, Yiming Sun, Mingyu Yan, Xiaochun Ye (Inst. of Computing Tech, CAS)
- Scaling Triangle Counting and K-Truss on the UpDown Architecture
- Jiya Su, Alexander Fell, Andronicus Rajasukumar (Univ. of Chicago), David F. Gleich (Purdue Univ.), Andrew A. Chien (Univ. of Chicago)
- PRISM: Practical In-Memory Acceleration for Subgraph Matching at Scale
- Deting Chen, Yu Huang, Yi Huang, Binbin Lin, Yi Zhang, Long Zheng, Xiaofei Liao, Hai Jin (Huazhong Univ. of Science and Tech.)
- BUG: Balanced DFS-Based Subgraph Matching with a Reuse Strategy on GPUs
- Zicang Xu, Lei Zou (Peking Univ.)
- InfraredGP: Efficient Graph Partitioning via Spectral Graph Neural Networks with Negative Corrections
- Meng Qin (Pengcheng Laboratory), Weihua Li (Beihang Univ.), Jinqiang Cui (Pengcheng Laboratory), Sen Pei (Columbia Univ.)
Thursday, September 18
4-K: Keynote Session (10:30-11:00)
Co-Chairs: J. Kepner & A. Reuther
- Keynote Talk: SPACE MICE: The Next Generation Data Systems
- Joshua Patterson (NVIDIA)
4-1: High Performance Computing Session (11:00-12:15)
Co-Chairs: D. Cousins & H. Sadasivan
- AGCRS: An Adaptive Generalized Storage Scheme for Large Sparse Tensors
- Md Mehrab Hossain Opi, K. M. Azharul Hasan (Khulna Univ. of Engr. and Tech.)
- Accelerating Supercomputing: AI-Hardware-Driven Innovation for Speed and Efficiency
- Jack Dongarra (Univ. of Tennessee), John Gunnels, Harun Bayraktar, Azzam Haidar, Dan Ernst (NVIDIA)
- On the Landscape of Scientific Computing Libraries in Python
- Niteya Shah (Virginia Tech), Pi-Yeuh Chuang (Argonne National Laboratory), Paul Sathre, Wu-chun Feng (Virginia Tech)
- Load Imbalance in HPC Applications: Improved Profiling and New Ways to Use Wasted Cycles
- Shining Yang, Xiteng Yao (Boston Univ.), Grace Nansamba, Amr Akmal Abouelmagd, Anthony Skjellum (Tennessee Tech), Martin Herbordt (Boston Univ.)
4-P1 (12:15-13:15): AI/ML/GenAI Poster Session
Chair(s)/Host(s): R. Lafuente-Mercado
Mitigation of Applied Load Using Machine Learning for Adaptive Motor Control
- Kourosh Rahnamai, Jacob Rollins, Luke Moisan, Ryan Kayfus (Western New England Univ.)
An Analysis of the New EU AI Act and A Proposed Standardization Framework for Machine Learning Fairness
Mike H.M. Teodorescu, Yongxu Sun, Haren N. Bhatia (Univ. of Washington), Christos Makridis (Univ. of Nicosia)
Reimagining Tacit Knowledge Extraction and Summarization by Inferencing Large Language Model in a High-Performance Computing Platform
- Rahul Hardikar, Saurabh Barve, Shampa Sarkar, Revati Kulkarni (TCS)
Data-Driven Dynamic Algorithm Dispatch with Large Language Models [Outstanding Short Paper Award]
- Rushil Shah, Emmanuel Lujan, Rabab Alomairy, Alan Edelman (MIT)
4-2: High Performance Computing Session (12:30-13:45)
Co-Chairs: D. Cousins & L. Zaidenberg
- Easy Acceleration with Distributed Arrays
- Jeremy Kepner, Chanup Byun, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel Burrill, Vijay Gadepally, Ryan Haney, Michael Houle, Matthew Hubbell, Hayden Jananthan,Michael Jones, Piotr Luszczek, Lauren Milechin, Guillermo Morales, Julie Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Peter Michaleas (MIT)
- Performance Evaluation of LAPACK Using SVE Optimized BLAS Kernels
- Aniket P. Garade, Sushil Singh, Vishal Rayala, Deepika H. V., Haribabu P., S. A. Kumar, S. D. Sudarsan (C-DAC)
- Predicting HPC Job Run time with Realistic Data Using Application Input Parameters
- Kenneth Lamar (Univ. of Central Florida), Benjamin A. Allan, M. Scot Swan, James M. Brandt (SNL), Damian Dechev (Univ. of Central Florida)
- Generalized Methodology for Determining Numerical Features of Hardware Floating-Point Matrix Multipliers: Part I
- Faizan A. Khattak, Mantas Mikaitis (Leeds Univ.)
- Balancing Performance and Productivity: A Comparative Study of Apache Arrow vs. MPI
- Ritvik Prabhu, Wu-chun Feng (Virginia Tech)
4-P2 (13:45-14:45): AI/ML/GenAI Poster Session
Chair(s)/Host(s): S.Mohindra
Centralized vs. Decentralized Security for Space AI Systems? A New Look
- Noam Schmitt, Marc Lacoste (Orange)
- Improving Uncertainty Based Dataset Pruning with Density Estimation for Noisy Edge Environments
- Mike Soricelli, Yuchou Chang (UMass Dartmouth), Christopher J Hixenbaugh (NUWC)
Virtual Benchmarking for HPC Systems Using ExaDigiT and Calculon
- Srishti Kalepu (Georgia Inst. of Tech.), Wesley Brewer, Matthias Maiterth (ORNL), Richard Vuduc (Georgia Inst. of Tech.)
- An Improved Izhikevich Neuron Model with Time Delay and Adaptive Parameters for High-Performance Spiking Neural Networks
- Alissa Kane (UMass Dartmouth), Felipe Marcelino, Anton Spirkin (NUWC), Yuchou Chang (UMass Dartmouth)
- Adaptive Policy Synchronization for Scalable Reinforcement Learning
- Rodney Lafuente Mercado (MIT Lincoln Laboratory)
- LLMs in Crisis Triage: Benchmarking Zero-Shot Classification of Social Media
- Emma L. McDaniel, Samuel Scheele, Jeffrey Liu (MIT Lincoln Laboratory)
4-3: High Performance Computing Session (14:15-15:30)
Co-Chairs: D. Cousins & P. Moniticciolo
- Leveraging Caliper and Benchpark to Analyze MPI Communication Patterns: Insights from AMG2023, Kripke, and Laghos
- Grace Nansamba, Evelyn Namugwanya (Tennessee Tech), David Boehme, Dewi Yokelson (LLNL), Riley Shipley (Tennessee Tech), Derek Schafer (Univ. of New Mexico), Michael McKinsey, Olga Pearce (LLNL), Anthony Skjellum (Tennessee Tech)
- Employing High-Performance PETSc Network Simulation for Business Profit Analysis
- Abu Asaduzzaman, Nowshin Nawal (Wichita State Univ.)
- Performance Analysis of Inline Compression in pySDC [Outstanding Paper Award]
- Emily Lattanzio (Clemson Univ.), Sansriti Ranjan (Siemens EDA), Robert Underwood (Argonne National Laboratory), Thomas Baumann, Robert Speck (Jülich Supercomputing Centre), Jon C. Calhoun (Clemson Univ.)
- The NorthPole Validator: A Cycle-Accurate Simulator for HW/SW Codesign of a Prescheduled Neural Inference Accelerator
- Alexander Andreopoulos, Michael V. Debole, Jeffrey A. Kusnitz, Nathaniel J. McClatchey, Tapan K. Nayak, Daniel F. Smith, Brian Taba, Filipp Akopyan, Rathinakumar Appuswamy, John V. Arthur, Andrew S. Cassidy, Pallab Datta, Carlos Ortega Otero, William P. Risk, Jun Sawada, Myron D. Flickner, Dharmendra S. Modha (IBM)
4-4: AI/ML/GenAI Session (15:45-17:15)
Co-Chairs: S. Mohindra & P. Moniticciolo
- Evaluating Efficiency and Novelty of LLM-Generated Code for Graph Analysis [Outstanding Student Paper Award]
- Atieh Barati Nia, Mohammad Dindoost, David A. Bader (New Jersey Inst. of Tech.)
- Enhancing Sentiment Classification of E-commerce Reviews for Actionable Insights using LLMs and NLP
- Kevin Power, Peter Harding, Jose Lopez, Alex Carroll, Elenna Dugundji (MIT)
- DNN-Driven Task Scheduling for High Performance Edge-Cloud Heterogeneous Systems
- Md Raihan Uddin, Abu Asaduzzaman, Fairuz Nawar, Christian Thompson (Wichita State Univ.)
- Automating Harmonized System (HS) Code Classification from Unstructured Shipping Manifests using Large Language Models
- Thomas Koch, Kevin Power (MIT)
- A Time-Aware Sliding Window-Based Hotel Recommendation Framework Using Multi-Stage BERT-MRC
- Md. Nazirul Hasan Shawon, K. M. Azharul Hasan (Khulna Univ. of Engr. and Tech.)
- BanglaDocAtlas: A Multi-Class Annotated Dataset for Complex Bangla Document Layout Analysis
- Md Safayat Hossain, Jannatul Ferdous (United Intl. Univ.), Md Raihan Uddin (Wichita State Univ.), K. M. A. Hossain, M. I. Ahmed, M. A. Rahman (United Intl. Univ.), Asif Sushmit (Bengali.AI), Farig Sadeque, Swakkhar Shatabda (BRAC University), Abu Asaduzzaman (Wichita State Univ.)
4-5: GenAI Opportunities & AI Challenges Session (17:30-19:30)
Organizers: V. Gadepally, D. Burrill & C. Prothmann
- MI300A vs H100, LLM Concerns for Applications at NRL, and MCP Testing
- William “Connor” Horne (NRL)
Trace Replay Simulation of MIT SuperCloud Dataset for Studying Optimal Policies for Sustainability [Outstanding Short Paper Award]
- Wesley Brewer, Matthias Maiterth (ORNL), Damien Fay (HPE)
- Introduction: Advancing AI Challenges for the United States Department of the Air Force
Christian Prothmann, Vijay Gadepally, Jeremy Kepner, Koley Borchard, Luca Carlone, Zachary Folcik,J. Daniel Griffith, Michael Houle, Jonathan P. How, Nathan Hughes, Ifueko Igbinedion, Hayden Jananthan, Tejas Jayashankar, Michael Jones, Sertac Karaman, Binoy G. Kurien, Alejandro Lancho, Giovanni Lavezzi, Gary C. F. Lee, Charles E. Leiserson, Richard Linares, Lindsey McEvoy, Peter Michaleas, Chasen Milner, Alex Pentland, Yury Polyanskiy, Jovan Popovich, Jeffrey Price, Tim W. Reid, Stephanie Riley, Siddharth Samsi, Peter Saunders, Olga Simek, Mark S. Veillette, Amir Weiss, Gregory W. Wornell, Daniela Rus, Scott T. Ruppel (MIT)
- Data-Driven Radio-Frequency Signal Separation Challenge
- Yury Polyanskiy (MIT)
- Tornado Network Challenge
- Mark Veillette (MIT Lincoln Laboratory), Peter Saunders (USAF)
- AI Innovation in Space Challenge
- Giovanni Lavezzi, Richard Linares (MIT)
- Evaluating Long-Context LLM Architectures with Controlled Synthetic Benchmarks Challenge
- Olga Simek (MIT Lincoln Laboratory)
- Predicting LLM Inference Server Request Capacity
- Daniel J. Burrill, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alexander Bonn, Chansup Byun, Michael Houle, Matthew Hubbell, Michael Jones, Piotr Luszczek, Peter Michaleas, Guillermo Morales, Julie Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee (MIT Lincoln Laboratory), Vijay Gadepally (MIT)
Friday, September 19
5-K: Keynote Session (10:30-11:00)
Co-Chairs: J. Kepner & A. Reuther
- Keynote Talk: Performance Engineering with MPI
- Bill Gropp (NCSA)
5-1: Mixed Precision Session (11:00-12:15)
Co-Chairs: P. Luszczek & C. Byun
- Invited Talk: Is Mixed Precision Computing really the Top Priority?
- Hartwig Anzt (TU Munich)
- Invited Talk: Reducing Numerical Precision Requirements in Quantum Chemistry Calculations
- William Dawson (RIKEN)
- Invited Talk: Why Exact Dot Products Obviate the Need for Mixed Precision
- John Gustafson (ASU)
- Invited Talk: Mixed Feelings about Mixed Precision
- Hatem Ltaief (KAUST)
- Invited Talk: Mixed Precision or Mixed Storage? User-Guided Compiler Transformations to Change the Data Layout On-the-Fly
- Tobias Weinzierl (Durham Univ.)
5-2: Mixed Precision Session (12:30-13:45)
Co-Chairs: P. Luszczek & R. Muri
- Invited Talk: Emulation Method for Matrix Multiplication
- Katsuhisa Ozaki (Shibaura Inst. of Tech.)
- A variable-precision implementation of the ADER-DG algorithm
- Marc Marot-Lassauzaie, Michael Bader (Tech. Univ. Munich)
- GPU-Accelerated, Mixed Precision GMRES(m) with Varied Restarts [Outstanding Student Paper Award]
- Abir Haque, Suzanne Shontz, Xuemin Tu (Univ. of Kansas)
- Performance and Numerical Aspects of Decompositional Factorizations with FP64 Floating-Point Emulation in INT8
- Piotr Luszczek, Vijay Gadepally, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel J. Burrill, Chansup Byun, Michael Houle, Matthew Hubbell, Michael Jones, Peter Michaleas, Guillermo Morales, Julia Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Jeremy Kepner (MIT Lincoln Laboratory)
- Adaptive Spectral Block Floating Point for Discontinuous Galerkin Methods
- Shivam Sundriyal, Markus Büttner (Univ. of Bayreuth), Christoph Alt, Tobias Kenter (Paderborn Univ.), Vadym Aizinger (Univ. of Bayreuth)
5-P2 (13:45-14:45): Embedded & GPU Poster Session
Chair(s)/Host(s): S. Shankar
Decision Support for Sustainable Agriculture Using I2C Sensors
- Anita Esmaeilian, Kishwar Ahmed (Univ. of Toledo)
ORB Feature Extraction on Embedded Platforms: A Heterogeneous CPU-GPU-PVA Approach
- Hamid Moghadaspour (Univ. of Coimbra), Nuno Neves (Univ. of Lisbon), Oscar Ferraz, Gabriel Falcao (Univ. of Coimbra)
- Accelerating Temporal Triangle Counting and Betweenness Centrality on GPUs
- Tuteja Trimansingh Parvindersingh, Venkata Kalyan Tavva (IIT Ropar), Subhasis Banerjee, Chiranjib Sur (Shell India Markets Pvt. Ltd.)
- SCoDa: Scalable Community Detection in Data Streams
- Akanksha Dwivedi, Prashant Srivastav, Dip Sankar Banerjee (IIT Jodhpur)
- CTAM Tool for Hyperscaler Qualification
- Tommy Yan, Rajat Madhusudan, Vani Pulendra, Anna Mary Mathew (Silicon Cloud HIE)
5-3: Cyber Analysis and Secure Computing Session (14:15-15:30)
Co-Chairs: D. Cousins & R. Vuduc
- Invited Talk: Private Database Analytics with PAC Privacy
- Srini Devadas (MIT)
- Accelerating Multi-Party Computation Using Heterogeneous Systems [Outstanding Student Paper Award]
- Xiteng Yao, Shining Yang, Mayank Varia, Martin Herbordt (Boston Univ.)
- Optimizing Local Computation in Secure Matrix Multiplication for Outsourced Neural Networks [Outstanding Paper Award]
- J. Parker Diamond, Andrea Lin, R. Nicholas Cunningham (MIT Lincoln Laboratory), Soamar Homsi (AFRL), John Darby Mitchell, Aseemit Pandey, Emily Shen (MIT Lincoln Laboratory)
- A Framework For The Iterative Solution of Sparse Linear Systems on Hybrid Architectures Using Homomorphic Encryption
- Lior Horesh, Vasileios Kalantzis, Barry M. Trager, Shashanka Ubaru (IBM Research)
- Secure Virtual Network Embedding Through Fully Homomorphic Encryption
- David Bruce Cousins, Carlo Pascoe (Duality Tech.), Erik Kline (USC)
5-4: Embedded Computing Session (15:45-17:15)
Co-Chairs: D. Cousins & E. Schnetzer
- Invited Talk: AI and Biodiversity
- Sara Beery (MIT)
- On the Adaptation of Mixed-Radix Fast Fourier Transform for Resource-Constrained Environments [Outstanding Student Paper Award]
- Atharva Gondhalekar, Paul Sathre, Wu-chun Feng (Virginia Tech)
- Comparative Analysis of RISC-V Softcore and Hardcore Processors for Space Computing [Outstanding Student Paper Award]
- Ni Nyoman Dhinar Gayatri, Alan D. George (Univ. of Pittsburgh)
- TRACER Software Switching at Waveform Timescales
- Connor Imes, Dong In D. Kang, Matthew French, John Paul Walters (USC Information Sciences Institute)
- Hardware-Accelerated Transformer Framework for Real-Time Battery SoH Estimation
- Talha Coskun (Univ. of Illinois Urbana-Champaign), Hiruna Vishwamith (Univ. of Moratuwa), Murat Isik (Stanford Univ.), I. Can Dikmen (Istinye Univ.)
- A Qualifiable GPU Sharing Approach for AI Workloads in Critical Systems
- Marc Solé i Bonet, Jannis Wolf (Barcelona Supercomputing Ctr.), Aridane Álvarez Suárez (fentISS), Leonidas Kosmidis (Barcelona Supercomputing Ctr.)
5-5: AI for Performance Engineering Session (17:30-19:30)
Organizers: H. Nguyen & D. Burrill
- UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC [Outstanding Student Paper Award]
- Tomer Bitan (Technion), Tal Kadosh (Ben-Gurion Univ. / IAEC), Erel Kaplan, Shira Meiri (Technion), Le Chen (Argonne NL), Peter Morales, Niranjan Hasabnis (Code Metal), Gal Oren (Stanford Univ.)
- Web-Based Intelligent Decision Support System for Real-Time Toll Plaza Management and AI-Driven Operational Optimization
- Pattarapon Klaykul, Wilaiporn Lee, Kanabadee Srisomboon, Luepol Pipanmekaporn, Akara Prayote (King Mongkut’s Univ.)
- CRAMP: Categorizing Classifiers and Regressors for Scalable Parallelism on Distributed and Multicore Systems
- Baidya Nath Saha, Pavan Sarvaiya, Wali Mohammad Abdullah, Md. Morshedul Islam (Concordia Univ. of Edmonton)
- RAILS: Retrieval-Augmented Intelligence for Learning Software Development
- Wali Mohammad Abdullah, Md. Morshedul Islam, Devraj Parmar, Happy Hasmukhbhai Patel, Sindhuja Prabhakaran, Baidya Saha (Concordia Univ. of Edmonton)
- Towards Automated Reasoning Chains for Verification of LLM-Generated Scientific Code
- Quentin Oschatz, Naifeng Zhang (Carnegie Mellon Univ.), Mike Franusich (SpiralGen), Franz Franchetti (Carnegie Mellon Univ.)
- P4OMP: Retrieval-Augmented Prompting for OpenMP Parallelism in Serial Code
- Wali Mohammad Abdullah (Concordia Univ. of Edmonton), Azmain Kabir (Univ. of Manitoba)
- Towards -OmL: A Deep Learning Based Approach to Outperform Compiler Defaults
- Hafsah Shahzad (Boston Univ.), Ahmed Sanaullah, Sanjay Arora, Ulrich Drepper (Red Hat), Martin Herbordt (Boston Univ.)