29th Annual
IEEE High Performance Extreme Computing Virtual Conference
15 - 19 September 2025

HPEC 2025 Agenda

Already registered?  Join the virtual conference here: https://www.engagez.net/HPEC2025

All times are EDT (UTC/GMT -04 hours)

Speaker/Presenting Author in Italics

Day Monday Tuesday Wednesday Thursday Friday
10:30-11:00am Session 1-K: Keynote Session 2-K: Keynote Session 3-K: Keynote Session 4-K: Keynote Session 5-K: Keynote
11:00am-12:15pm Session 1-1: Advanced Processor Architectures Session 2-1: Bridging Quantum and HPC Session 3-1: Scaling Research Computing Education Session 4-1: High Performance Computing Session 5-1: Mixed Precision
12:15-12:30pm Break Poster Session 1-P1 (12:15-13:15): New Application Frontiers Break Poster Session 2-P1 (12:15-13:15): Graph AI & Sparse Break Tutorial Session 3-T (12:15-15:45): Spiral Tutorial Break Poster Session 4-P1 (12:15-13:15): AI/ML/GenAI Break
12:30-1:45pm Session 1-2: Neuromorphic AI Session 2-2: Case Studies, Benchmarking, and Tools Session 3-2: AI in the World Session 4-2: High Performance Computing Session 5-2: Mixed Precision
1:45-2:15pm Break Poster Session 1-P2 (13:45-14:45): High Performance Data Analysis Break Break Poster Session 3-P2 (13:45-14:45): High Performance Computing Break Poster Session 4-P2 (13:45-14:45): AI/ML/GenAI Break Poster Session 5-P2 (13:45-14:45): Embedded & GPU
2:15-3:30pm Session 1-3: ASIC and FPGA Advances Session 2-3: Quantum and Non-Deterministic Computing Session 3-3: Graph AI & Sparse Session 4-3: High Performance Computing Session 5-3: Cyber Analysis and Secure Computing
3:30-3:45pm Break Break Break Break Break
3:45-5:00pm Session 1-4: FastCode Session 2-4: Graph AI & Sparse Session 3-4: Graph AI and Sparse Session 4-4: AI/ML/GenAI Session 5-4: Embedded Computing
5:00-5:30pm Break Break Break Break Break
5:30-7:30pm Session 1-5: BRAINS – Building Resilience through Artificial Intelligence for Networked Systems Session 2-5: GraphBLAS BoF Session 3-5: Graph Challenge Session 4-5: GenAI Opportunities & AI Challenges Session 5-5: AI for Performance Engineering

Monday, September 15

 

1-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther  
Keynote Talk: Enabling Advances in OceanAI
Nick Rotker (MITRE)
 

1-1: Advanced Processor Architectures Session (11:00-12:15)

Co-Chairs: P. Luszczek & C. Byun  
Invited Talk:
Michael Foertsch (Q.ANT)
Encoded Time-Series Model Training with UNet Running on Wafer Scale Engine
Vyacheslav Romanov (NETL)
Lincoln AI Computing Survey (LAICS) and Trends
Albert Reuther, Peter Michaleas, Michael Jones, Vijay Gadepally, Jeremy Kepner (MIT Lincoln Laboratory Supercomputing Center)
System-Level Performance Modeling of Photonic In-Memory Computing
Jebacyril Arockiaraj, Sasindu Wijeratne (USC), Sugeet Sunder (USC Information Sciences Institute), Md Abdullah-Al Kaiser, Akhilesh Jaiswal (Univ. of Wisconsin), Ajey P Jacob (USC Information Sciences Institute), Viktor Prasanna (USC)
 

1-P1 (12:15-13:15): New Application Frontiers Poster Session

Chair(s)/Host(s): K. Keville
Multi-Stage Stochastic Programming for Heavy-Duty Electric Truck Routing Under Public Charging Congestion Uncertainty
Ziyan Li, Nikolay Aristov, Antoine Germain, Elenna R. Dugundji (MIT)
Enabling Heterogeneous Performance Analysis for Scientific Workloads [Outstanding Short Paper Award]
Maksymilian Graczyk (CERN), Vincent Desbiolles (HES-SO), Stefan Roiser (CERN), Andrea Guerrieri (HES-SO and EPFL)
Weighted Histogram Matching for Improved Automated Synthetic Aperture Radar (SAR)-optical Image Registration
Kerri Prinos, Aditi Mungale, Dr. Karen Gettings, Dr. Tair Akhmejanov, Devanshu Mehta (MIT Lincoln Laboratory)
Towards a Framework for Etymological Analysis of Literary and AI Works
Mike Teodorescu (Univ. of Washington), Cecilia Speranța Bolea (IIT), Horia Teodorescu (Technical Univ. Iasi)
An Algorithmically Determined Exchange-Correlation (ADEXC) Method for Electronic Structure Calculations
Ivan Williams, Eric Polizzi (UMass Amherst)
Scaling Performance of Large Language Model Pretraining [Best Short Paper Award]
Alexander Interrante-Grant, Carla Varela-Rosa, Suhaas Narayan, Chris Connelly, Albert Reuther (MIT Lincoln Laboratory)
Metadata Guided Pose Estimation for 3D Reconstruction
William Farthing (Emory Univ.), Compton Ross (Univ. of Mississippi), Jaired Collins, Jing-Ru C. Cheng (U.S. Army Corp of Engineers)
Neuromorphic Processor Employing FPGA Technology with universal interconnections
Pracheta Harlikar, Abdel-Hameed Badawy (New Mexico State Univ.), Prasanna Date (ORNL)
 

1-2: Neuromorphic AI Session (12:30-13:45)

Co-Chairs: M. Barnell & H. Nguyen  
Artificial Intelligence Performance and Radiation Effects in Neuromorphic NorthPole Hardware
Victor M. Vergara (BlueHalo-AFRL), Francisco O. Viramontes, Amanda E. Romero (COSMIAC Research Center), Matthew E. Spear, Evan T. Kain, Windy S. Slater, Heather M. Quinn, Qing Wu (AFRL)
Physics-Informed Neural Networks for Low-Power Real-Time Edge Biosensing Applications
Soheli Farhana (Harvard Univ.)
Analysis and Optimization of Spiking Neural Network Simulations on GPUs
Doğu Kocatepe, Işıl Öz (Izmir Institute of Technology)
Exploring Neuromorphic Computing with Loihi-2 for High-Performance CFD Simulations
Talha Coskun (Univ. of Illinois Urbana-Champaign), Hiruna Vishwamith (Univ. of Moratuwa), Murat Isik (Stanford Univ.), I. Can Dikmen (Istinye Univ.)
Neural-Inspired Enhancing Spiking Graph Convolutional Networks
Fernando Vera Buschmann, Horacio Roststein, Vincent Oria (New Jersey Inst. of Tech.)
 

1-P2 (13:45-14:45): High Performance Data Analysis Poster Session

Chair(s)/Host(s): K. Keville  
Accelerating Push-Relabel Algorithm on GPU via Two-Level Parallelism Paradigm and Efficient CSR Designs
Chou Ying Hsieh, Po-Chieh Lin, Sy-Yen Kuo (National Taiwan Univ.)
Improving Statistical Characterization of Data Tensors with the Generalized Canonical Polyadic Tensor Decomposition
Matthew Merris, Tim Andersen (Boise State Univ.)
Comparative Analysis of Classical and Deep Learning Features for Texture Image Classification
Murad Hossen (Univ. of Houston)
Aspects of HPC within Conformal Recommendation Systems: Ideas and Challenges
Stanisław M. S. Halkiewicz (AGH Univ. of Cracow), Maciej Kuczyński (Czestochowa Univ. of Tech.)
Enhancing the Real-Time Solutions of Parametric Linear Systems on a GPU through Hybrid Coarse-Grained Transprecision Computing
Hamid Noori, Hans Vandierendonck, Roger Woods (Queen’s Univ. Belfast), Nick Polydorides (Univ. of Edinburgh)
AAPA: An Archetype-Aware Predictive Autoscaler with Uncertainty Quantification for Serverless Workloads on Kubernetes
Guilin Zhang (George Washington Univ.), Srinivas Vippagunta, Raghavendra Nandagopal, Suchitra Raman, Jeff Xu, Marcus Pfeiffer, Shreeshankar Chatterjee (Workday), Ziqi Tan, Wulan Guo (George Washington Univ.), Hailong Jiang (Youngstown State Univ.)
Stable Iterative Solvers for Ill-Conditioned Linear Systems and Least Squares
Vasileios Kalantzis, Mark S. Squillante, Chai Wah Wu (IBM Research)
 

1-3: ASIC and FPGA Advances Session (14:15-15:30)

Co-Chairs: L. Zaidenberg & K. Thurmer  
NTT-SAA: Exploring NTT Acceleration with 2-D Systolic Array Architecture on FPGAs
Ashwajit Singh (IIT Bombay), Zhihan Xu, Viktor K. Prasanna (USC)
Evaluating AMD-Xilinx Frameworks for Deep-Learning Acceleration on Versal [Outstanding Student Paper Award]
Peter Drum, Alan D. George (NSF SHREC)
Fast FPGA-Based Implementation of the QP-Dyn Stream Cipher using High-Level Synthesis
Paolo Palazzari (ENEA), Luigi Accardi (Volterra Univ.), Antonio Mastrandrea, Pasquale Tommasino, Alessandro Trifiletti (Sapienza Univ.)
A Scalable Code Generation Flow for Heterogeneous Parallel RTL Simulation using MLIR
Jie Tong, Zhengxiong Li, Umit Yusuf Ogras, Tsung-Wei Huang (Univ. of Wisconsin)
Accelerating Dynamic Image Graph Construction on FPGA for Vision GNNs
Anvitha Ramachandran, Dhruv Parikh, Viktor Prasanna (USC)
 

1-4: FastCode Session (15:45-17:15)

Co-Chairs: Bruce Hoppe  
Invited Talk: Fastcode: An Open-Source Community for Making Software Performance Engineering Easy and Fun
Bruce Hoppe (MIT)
Invited Talk: Taskflow — A General-Purpose Task-Parallel Programming System
Tsung-Wei Huang (Univ. of Wisconsin)
Invited Talk: OpenCilk — A Modular and Extensible Software Infrastructure for Fast Task-Parallel Code
I-Ting Angelina Lee (WUSTL)
Invited Talk: Multithreaded Parallel Python Through OpenMP Support in Numba
Tim Mattson (Human Learning Group)
 

1-5: BRAINS – Building Resilience through Artificial Intelligence for Networked Systems Session (17:30-19:30)

Organizers: T. Hardjono & S. Pisharody  
Invited Talk: Developing Trust in the Supply Chain
Guy Fedorkow (Juniper/HPE)
Invited Talk: Using AI to Achieve Peak Supply Chain Health
Karyl Fowler (Tradeverifyd)
Invited Talk: Supply Chain Transparency as Entailment for Device Attestation
Ned Smith (Intel)
Towards an Algorithm-based Approach for Soft Error Tolerance using Interval Arithmetic [Best Paper Award]
Larry Tang, Varun Kumar, Matt Ngaw, Siddharth Singh, Devdutt Nadkarni, Lohith Tummala, Ken Mai, Franz Franchetti (Carnegie Mellon Univ.)
Detecting Expert-Written Comments in Stack Exchange
Himani Musku (Carnegie Mellon Univ.), Alea Ritchie (Stanford Univ.), Nour Jedidi, Rohan Leekha, Courtland VanDam (MIT Lincoln Laboratory)
Strategic Cyber Defense via RL-Guided Combinatorial Auctions
Mai Pham, Vikrant Vaze, Peter Chin (Dartmouth Coll.)
AI-Powered Network Energy Optimization Machine Learning Approaches to Reducing Network Power Consumption
Het Mehta (Cisco Systems), Sambu Patach Arrojula (Samsung)
   

Tuesday, September 16

 

2-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther  
Keynote Talk: Effective Human-Machine Partnerships in High Stakes Settings
Julie Shah (MIT)
 

2-1: Bridging Quantum and HPC Session (11:00-12:15)

Co-Chairs: Devesh Tiwari  
Invited Talk: Tightening the Integration of Quantum Computers with HPC Systems
Travis Humble (ORNL)
Invited Talk: Enabling Quantum Utility through a System Toolkit
Tirthak Patel (Rice Univ.)
Invited Talk: CUDA-Q: Defining a Tightly-Integrated System Software Stack for Quantum-Accelerated HPC
Alex McCaskey (NVIDIA)
 

2-P1 (12:15-13:15): Graph AI & Sparse Poster Session

Chair(s)/Host(s): K. Cain  
Cyber Orbits of Large Scale Network Traffic
Jeremy Kepner, Hayden Jananthan, Chasen Milner, Michael Houle, Michael Jones, Peter Michaleas, Alex Pentland (MIT)
Accelerating AI Development with Cyber Arenas
William Cashman, Chasen Milner (USAF), Michael Houle, Michael Jones, Hayden Jananthan, Jeremy Kepner, Peter Michaleas, Alex Pentland (MIT)
Optimizing Sparse Matrix-Vector Multiplication on GPUs using the Mathematics of Arrays
Stephen Thomas (Lehigh Univ.), Lenore Mullin (Univ. of Albany)
Sampling to Scale: Performance Trade-offs in Approximate Triangle and Square Counting
Shubhashish Kar, Shaikh Arifuzzaman (UNLV)
Spectral Sparsification of Edges for Efficient Clique Counting
Aaron Schindler (Univ. of Utah), Hari Sundar (Tufts Univ.)
Degree Matrix Comparison for Graph Alignment
Ashley Wang, Peter Chin (Dartmouth Coll.)
 

2-2: Case Studies, Benchmarking, and Tools Session (12:30-13:45)

Co-Chairs: X. Sun & J. Ghanem  
Invited Talk: Rigorous Mathematics in Data-driven Methods for Safe Autonomy
Chuchu Fan (MIT)
AOCL-CST – A new CPU Stress Test library for AMD CPUs
S. Biplab Raut (AMD)
Performance–Energy Characterization of ML Inference on Heterogeneous Edge AI Platforms
Palash Kohli (BITS Pilani), Rakshith Jayanth, Neelesh Gupta, Haoyang Fan, Viktor Prasanna (USC)
A Fully Adaptive Radau Method for the Efficient Solution of Stiff Ordinary Differential Equations at Low Tolerances
Shreyas Ekanathan (Lexington High School), Oscar Smith (JuliaHub), Christopher Rackauckas (MIT)
Evaluation of Habitat Robotics using Large Language Models
William Li, Lei Hamilton, Kaise Al-natour, Sanjeev Mohindra (MIT Lincoln Laboratory)
 

2-3: Quantum and Non-Deterministic Computing Session (14:15-15:30)

Co-Chairs: B. Sroka & I. DeTore  
SYCL for Performance Portability: Application Experience with Coupled Cluster Formalism in Quantum Chemistry on Exascale Systems [Outstanding Paper Award]
Abhishek Bagusetty (Argonne National Laboratory), Ajay Panyala (PNNL), Alvaro Vazquez Mayagoitia (Argonne National Laboratory), John K. Holmen (ORNL), Kevin Harms (Argonne National Laboratory)
Implementation of Tensor Network Simulation TN-Sim under NWQ-Sim
Aaron C. Hoyt, Jonathan S. Bersson, Sean Garner (Univ. of Washington), Chenxu Liu, Ang Li (PNNL)
Partition-based Surface Code Compilation
Hanjing Xu (Purdue Univ.), Xiaoyuan Liu, Ankit Kulshrestha, Hayato Ushijima-Mwesigwa (Fujitsu Research of America)
QAOA Parameter Transferability for Maximum Independent Set using Graph Attention Networks
Hanjing Xu (Purdue Univ.), Xiaoyuan Liu (Fujitsu Research of America), Alex Pothen (Purdue Univ.), Ilya Safro (Univ. of Delaware)
A Hybrid Classical-Quantum Model for QSAR-Based Biodegradability Prediction
Batuhan Hangun, Oguz Altun (Yıldız Technical Univ.), Onder Eyecioglu (Bolu Abant İzzet Baysal Univ.)
 

2-4: Graph AI & Sparse Session (15:45-17:15)

Co-Chairs: X. Sun & J. Ghanem  
Scalable Graph Algorithms on Distributed UpDown Accelerators
Brian Wheatman, Andrew A. Chien (Univ. of Chicago)
2D Distributed Label Propagation on 400 GPUs
George M. Slota, Michael Mandulak, Ujwal Pandey, Anthony Fabius (RPI)
Accelerating Sparse Deep Learning via Multi-Layer Tensor Reordering and Partitioning
Gunduz Vehbi Demirci (Imagination Tech.), Cagatay Dikici (Wayve), Tim Atherton (Imagination Tech.)
Julia GraphBLAS with Nonblocking Execution [Outstanding Paper Award]
Pascal Costanza (Independent Researcher), Timothy G. Mattson (Univ. of Bristol), Raye Kimmerer (NERSC LBNL), Benjamin Brock (Intel)
A parallel push-relabel maximum flow algorithm in LAGraph and GraphBLAS
Darin A. Peries, Timothy A. Davis (Texas A&M Univ.)
GraphBLAS Mathematical Opportunities: Parallel Hypersparse, Matrix Based Graph Streaming, and Complex-Index Matrices
Hayden Jananthan, Jeremy Kepner, Michael Jones, Vijay Gadepally, Michael Houle, Peter Michaleas, Chasen Milner, Alex Pentland (MIT)
 

2-5: GraphBLAS BoF Session (17:30-19:30)

Organizers: T. Mattson, B. Brock & S. McMillan  
GraphBLAS APIs: The Math Spec
Tim Mattson (Human Learning Group)
Keynote – Postgres and GraphBLAS
Michel Pelletier (OneSparse)
SuiteSparse GraphBLAS and LAGraph: Progress Report and Future Direction
Tim Davis (Texas A&M)
   

Wednesday, September 17

 

3-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther  
Keynote Talk: Building AI That Users Trust
Ashley Conard (Microsoft)
 

3-1: Scaling Research Computing Education Session (11:00-12:15)

Co-Chairs: J. Mullen, L. Milechin & H. Jananthan  
Invited Talk: Research/Advanced Computing Roles: Skillfully Facing the Challenges
Robert M. Freeman, Jr. (Harvard University & CaRCC)
Invited Talk: Skill Inventories, What They Are and Why We Need Them
Weronika Filinger (Edinburgh Parallel Computing Center)
Invited Talk: Cataloguing the Training Material Landscape: Skills, Gaps, and Resources
Jeremy Cohen (Imperial Col.)
Invited Talk: Designing an Integrated Ecosystem to Develop Research Computing and Digital Technical Professionals
Julia Mullen (MIT LL)
 

3-2: AI in the World Session (12:30-13:45)

Co-Chairs: K. Gettings & M. Barnell  
Invited Talk: Learning-Guided Optimization for Mobility
Cathy Wu (MIT)
Sustainably Modeling a Sustainable Future Climate
Rabab Alomairy (MIT), Sameh Abdulah (KAUST), Qinglei Cao (St. Louis Univ.), Marc G. Genton, David E. Keyes, Hatem Ltaief (KAUST)
Predicting Ports Congestion by Utilizing State Space Models
Nikolay Aristov, Elenna R. Dugundji (MIT-CTL)
Scaling Regime-Aware Forecasting: Distributed Shifting Seasonal Matrix Factorization
Jacob Munson, Breschine Cummins (Montana State Univ.)
AIMS: An Adaptive Intelligent Multi-Objective Scheduler Powered by Digital Twins
Kyrian Adimora, Hongyang Sun (Univ. of Kansas)
 

3-P2 (13:45-14:45): High Performance Computing Poster Session

Chair(s)/Host(s): P. Luszczek  
When Structure is Silent: Opportunities for Algorithmic Dispatch in Linear Algebra
Emmanuel Lujan, Alan Edelman (MIT)
Performance Modeling of Heterogeneous Edge-Cloud Systems with Machine Learning
Md Raihan Uddin, Abu Asaduzzaman, Sonu Gangadhar Gowda (Wichita State Univ.)
An FFT-based Preconditioner for Conjugate Gradient Pressure Solvers in Complex Domains
Xin Kai Lee, Gregory LeClaire Wagner, Simone Silvestri, Raffaele Ferrari (MIT)
A Scalable Quantum Dynamical Approach for Calculating Collisional Molecular Properties
Prajwal Niraula, Laurent Wiesenfeld, Julien de Wit (MIT), Iouli Gordon, Robert Hargreaves (Harvard Univ.), Jeremy Kepner, Deborah Woods, Cooper Loughlin (MIT Lincoln Laboratory)
Scalable Bayesian Nonparametric Ensemble (BNE) for Spatio-temporal Air Pollution Predictions using High-Performance Computing
Vijay Kumar, Jaime Benavides (Brown University), Carlos Carrillo-Gallegos (Columbia Univ.), Gil Speyer (Arizona State Univ.), Marianthi-Anna Kioumourtzoglou (Brown University)
Investigating the Impact of Algorithms and Hardware on Machine Learning Models in HPC Systems
Christian C. Thompson, Abu Asaduzzaman, Md Raihan Uddin (Wichita State Univ.)
SARComp: High-Performance Algorithms for Onboard SAR from FFT Kernels to Matched Filtering
Maron Schlemon (German Aerospace Center), Martin Schulz (Tech. Univ. Munich), Rolf Scheiber (German Aerospace Center)
Streamed Multi-Format Sparse Matrix Vector Multiplication for FPGA
Spencer Smith, Richard Veras (Univ. of Oklahoma)
 

3-3: Graph AI & Sparse Session (14:15-15:30)

Co-Chairs: N. Pitsianis & C. Byun  
pdGRASS: A Fast Parallel Density-Aware Algorithm for Graph Spectral Sparsification [Best Student Paper Award]
Tiancheng Zhao (Georgia Inst. of Tech.), Zekun Yin, Huihai An (Shandong Univ.), Xiaoyu Yang (China Univ. of Petroleum-Beijing), Zhou Jin (Zhejiang Univ.), Jiasi Shen (HKUST), Helen Xu (Georgia Inst. of Tech.)
Differentiable Graph Centrality
Georgios Kollias, Vassilis Kalantzis (IBM Research)
Performance Analysis of the Parallel Shared-Memory Sparse Matrix-Vector Multiplication on Unstructured Matrices
Kobe Bergmans, Karl Meerbergen, Raf Vandebril (KU Leuven)
HiPerMotif: Novel Parallel Subgraph Isomorphism in Large-Scale Property Graphs
Mohammad Dindoost, Oliver Alvarado Rodriguez, Bartosz Bryg, Ioannis Koutis, David A. Bader (New Jersey Inst. of Tech.)
Characterization of Sparsity-aware Parallelization of Jaccard Similarity in Graph Datasets
Atharva Gondhalekar, Paul Sathre, Wu-chun Feng (Virginia Tech)
 

3-4: Graph AI and Sparse Session (15:45-17:15)

Co-Chairs: N. Pitsianis & C. Byun  
Enhancing Graph Partitioning with Reinforcement Learning-based Initialization
Chedi Morchdi (Texas A&M Univ.), Cheng-Hsiang Chiu, Wan Luan Lee, Tsung-Wei Huang (Univ. of Wisconsin), Yi Zhou (Texas A&M Univ.)
Benchmarking Deep Learning with Representative ONNX Subgraphs
Marika E. Schubert (Univ. of Pittsburgh), David Langerman (NSF SHREC), Evan W. Gretok, Ian Peitzsch, Calvin B. Gealy, Jefferson Boothe, Alan D. George (Univ. of Pittsburgh)
A Locality Sensitive Hashing Based Algorithm to Accelerate Neighborhood Search in Graph Neural Operators
Mariam Hassan, Sanmukh Kuppannagari (Case Western Reserve Univ.)
Power Iteration with Probabilistic Updates for Systems with Heterogeneous Performance
Soumyadip Ghosh, Lior Horesh, Vasileios Kalantzis, Georgios Kollias, Yingdong Lu, Tomasz Nowicki, Shashanka Ubaru (IBM Research)
Designing Parallel Algorithms for Community Detection using Arachne
Fuhuan Li, Zhihui Du, David Bader (New Jersey Inst. of Tech.)
GCN-Driven CUDA Parameter Optimization for Parallel Triangle Counting in Graphs
Hasan Serdar Arikan, Rakibul Hassan, Shubhashish Kar (Univ. of Nevada Las Vegas), Doru Thom Popovici (LBNL), Shaikh Arifuzzaman (Univ. of Nevada Las Vegas)
 

3-5: Graph Challenge Session (17:30-19:30)

Organizers: J. Kepner & A. Reuther  
Combining Performance and Productivity: Accelerating the Network Sensing Graph Challenge with GPUs and Commodity Data Science Software
Siddharth Samsi, Dan Campbell, Emanuel Scoullos, Oded Green (NVIDIA)
SANST: Sensing Anonymized Network via Sorted Triplets
Jianyu Wang, Wenzi Tang, Chenglong Shi, Zhe Zhang (Guangxi Univ.), Dan Chen (National Univ. Singapore), Miaojiang Chen, Wenjing Xiao (Guangxi Univ.)
Anonymized Network Sensing using C++26 std::execution on GPUs
Michael Mandulak (RPI), Sayan Ghosh, S M Ferdous, Mahantesh Halappanavar (PNNL), George Slota (RPI)
Geans: A GPU-accelerated Framework for Efficient End-to-End Anonymized Network Sensing
Jun Mai, Qinggang Wang, Yu Huang, Pengcheng Yao, Long Zheng, Xiaofei Liao, Hai Jin (Huazhong Univ. of Science and Tech.)
DBOS Network Sensing: A Web Services Approach to Collaborative Awareness
Sophia Lockton, Jeremy Kepner, Michael Stonebraker, Hayden Jananthan, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel Burrill, Chansup Byun (MIT), Timothy Davis (Texas A&M), Vijay Gadepally, Michael Houle, Matthew Hubbell, Michael Jones, Piotr Luszczek, Peter Michaleas, Lauren Milechin, Chasen Milner, Guillermo Morales, Julie Mullen (MIT), Michel Pelletier (OneSparse), Alex Poliakov (DBOS), Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Alex Pentland (MIT)
Interactive Trillion Packet Anonymized Network Analysis with the GraphBLAS
Chasen Milner (USAF), Michael Houle, Hayden Jananthan, Michael Jones, Jeremy Kepner, Peter Michaleas, Inna Voloshchuk, Alex Pentland (MIT)
Towards Efficient Sparse Deep Neural Network Inference via Multi-level Concurrency Orchestration
Ming Dun, Jie Zhou (Inst. of Computing Tech, CAS), Huawei Cao (UCAS), Shuhan Song, Yiming Sun, Mingyu Yan, Xiaochun Ye (Inst. of Computing Tech, CAS)
Scaling Triangle Counting and K-Truss on the UpDown Architecture
Jiya Su, Alexander Fell, Andronicus Rajasukumar (Univ. of Chicago), David F. Gleich (Purdue Univ.), Andrew A. Chien (Univ. of Chicago)
PRISM: Practical In-Memory Acceleration for Subgraph Matching at Scale
Deting Chen, Yu Huang, Yi Huang, Binbin Lin, Yi Zhang, Long Zheng, Xiaofei Liao, Hai Jin (Huazhong Univ. of Science and Tech.)
BUG: Balanced DFS-Based Subgraph Matching with a Reuse Strategy on GPUs
Zicang Xu, Lei Zou (Peking Univ.)
InfraredGP: Efficient Graph Partitioning via Spectral Graph Neural Networks with Negative Corrections
Meng Qin (Pengcheng Laboratory), Weihua Li (Beihang Univ.), Jinqiang Cui (Pengcheng Laboratory), Sen Pei (Columbia Univ.)
   

Thursday, September 18

 

4-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther  
Keynote Talk: SPACE MICE: The Next Generation Data Systems
Joshua Patterson (NVIDIA)
 

4-1: High Performance Computing Session (11:00-12:15)

Co-Chairs: D. Cousins & H. Sadasivan  
AGCRS: An Adaptive Generalized Storage Scheme for Large Sparse Tensors
Md Mehrab Hossain Opi, K. M. Azharul Hasan (Khulna Univ. of Engr. and Tech.)
Accelerating Supercomputing: AI-Hardware-Driven Innovation for Speed and Efficiency
Jack Dongarra (Univ. of Tennessee), John Gunnels, Harun Bayraktar, Azzam Haidar, Dan Ernst (NVIDIA)
On the Landscape of Scientific Computing Libraries in Python
Niteya Shah (Virginia Tech), Pi-Yeuh Chuang (Argonne National Laboratory), Paul Sathre, Wu-chun Feng (Virginia Tech)
Load Imbalance in HPC Applications: Improved Profiling and New Ways to Use Wasted Cycles
Shining Yang, Xiteng Yao (Boston Univ.), Grace Nansamba, Amr Akmal Abouelmagd, Anthony Skjellum (Tennessee Tech), Martin Herbordt (Boston Univ.)
 

4-P1 (12:15-13:15): AI/ML/GenAI Poster Session

Chair(s)/Host(s): R. Lafuente-Mercado  
Mitigation of Applied Load Using Machine Learning for Adaptive Motor Control
Kourosh Rahnamai, Jacob Rollins, Luke Moisan, Ryan Kayfus (Western New England Univ.)
An Analysis of the New EU AI Act and A Proposed Standardization Framework for Machine Learning Fairness
Mike H.M. Teodorescu, Yongxu Sun, Haren N. Bhatia (Univ. of Washington), Christos Makridis (Univ. of Nicosia)
Reimagining Tacit Knowledge Extraction and Summarization by Inferencing Large Language Model in a High-Performance Computing Platform
Rahul Hardikar, Saurabh Barve, Shampa Sarkar, Revati Kulkarni (TCS)
Data-Driven Dynamic Algorithm Dispatch with Large Language Models [Outstanding Short Paper Award]
Rushil Shah, Emmanuel Lujan, Rabab Alomairy, Alan Edelman (MIT)
 

4-2: High Performance Computing Session (12:30-13:45)

Co-Chairs: D. Cousins & L. Zaidenberg  
Easy Acceleration with Distributed Arrays
Jeremy Kepner, Chanup Byun, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel Burrill, Vijay Gadepally, Ryan Haney, Michael Houle, Matthew Hubbell, Hayden Jananthan,Michael Jones, Piotr Luszczek, Lauren Milechin, Guillermo Morales, Julie Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Peter Michaleas (MIT)
Performance Evaluation of LAPACK Using SVE Optimized BLAS Kernels
Aniket P. Garade, Sushil Singh, Vishal Rayala, Deepika H. V., Haribabu P., S. A. Kumar, S. D. Sudarsan (C-DAC)
Predicting HPC Job Run time with Realistic Data Using Application Input Parameters
Kenneth Lamar (Univ. of Central Florida), Benjamin A. Allan, M. Scot Swan, James M. Brandt (SNL), Damian Dechev (Univ. of Central Florida)
Generalized Methodology for Determining Numerical Features of Hardware Floating-Point Matrix Multipliers: Part I
Faizan A. Khattak, Mantas Mikaitis (Leeds Univ.)
Balancing Performance and Productivity: A Comparative Study of Apache Arrow vs. MPI
Ritvik Prabhu, Wu-chun Feng (Virginia Tech)
 

4-P2 (13:45-14:45): AI/ML/GenAI Poster Session

Chair(s)/Host(s): S.Mohindra  
Centralized vs. Decentralized Security for Space AI Systems? A New Look
Noam Schmitt, Marc Lacoste (Orange)
Improving Uncertainty Based Dataset Pruning with Density Estimation for Noisy Edge Environments
Mike Soricelli, Yuchou Chang (UMass Dartmouth), Christopher J Hixenbaugh (NUWC)
Virtual Benchmarking for HPC Systems Using ExaDigiT and Calculon
Srishti Kalepu (Georgia Inst. of Tech.), Wesley Brewer, Matthias Maiterth (ORNL), Richard Vuduc (Georgia Inst. of Tech.)
An Improved Izhikevich Neuron Model with Time Delay and Adaptive Parameters for High-Performance Spiking Neural Networks
Alissa Kane (UMass Dartmouth), Felipe Marcelino, Anton Spirkin (NUWC), Yuchou Chang (UMass Dartmouth)
Adaptive Policy Synchronization for Scalable Reinforcement Learning
Rodney Lafuente Mercado (MIT Lincoln Laboratory)
LLMs in Crisis Triage: Benchmarking Zero-Shot Classification of Social Media
Emma L. McDaniel, Samuel Scheele, Jeffrey Liu (MIT Lincoln Laboratory)
 

4-3: High Performance Computing Session (14:15-15:30)

Co-Chairs: D. Cousins & P. Moniticciolo  
Leveraging Caliper and Benchpark to Analyze MPI Communication Patterns: Insights from AMG2023, Kripke, and Laghos
Grace Nansamba, Evelyn Namugwanya (Tennessee Tech), David Boehme, Dewi Yokelson (LLNL), Riley Shipley (Tennessee Tech), Derek Schafer (Univ. of New Mexico), Michael McKinsey, Olga Pearce (LLNL), Anthony Skjellum (Tennessee Tech)
Employing High-Performance PETSc Network Simulation for Business Profit Analysis
Abu Asaduzzaman, Nowshin Nawal (Wichita State Univ.)
Performance Analysis of Inline Compression in pySDC [Outstanding Paper Award]
Emily Lattanzio (Clemson Univ.), Sansriti Ranjan (Siemens EDA), Robert Underwood (Argonne National Laboratory), Thomas Baumann, Robert Speck (Jülich Supercomputing Centre), Jon C. Calhoun (Clemson Univ.)
The NorthPole Validator: A Cycle-Accurate Simulator for HW/SW Codesign of a Prescheduled Neural Inference Accelerator
Alexander Andreopoulos, Michael V. Debole, Jeffrey A. Kusnitz, Nathaniel J. McClatchey, Tapan K. Nayak, Daniel F. Smith, Brian Taba, Filipp Akopyan, Rathinakumar Appuswamy, John V. Arthur, Andrew S. Cassidy, Pallab Datta, Carlos Ortega Otero, William P. Risk, Jun Sawada, Myron D. Flickner, Dharmendra S. Modha (IBM)
 

4-4: AI/ML/GenAI Session (15:45-17:15)

Co-Chairs: S. Mohindra & P. Moniticciolo  
Evaluating Efficiency and Novelty of LLM-Generated Code for Graph Analysis [Outstanding Student Paper Award]
Atieh Barati Nia, Mohammad Dindoost, David A. Bader (New Jersey Inst. of Tech.)
Enhancing Sentiment Classification of E-commerce Reviews for Actionable Insights using LLMs and NLP
Kevin Power, Peter Harding, Jose Lopez, Alex Carroll, Elenna Dugundji (MIT)
DNN-Driven Task Scheduling for High Performance Edge-Cloud Heterogeneous Systems
Md Raihan Uddin, Abu Asaduzzaman, Fairuz Nawar, Christian Thompson (Wichita State Univ.)
Automating Harmonized System (HS) Code Classification from Unstructured Shipping Manifests using Large Language Models
Thomas Koch, Kevin Power (MIT)
A Time-Aware Sliding Window-Based Hotel Recommendation Framework Using Multi-Stage BERT-MRC
Md. Nazirul Hasan Shawon, K. M. Azharul Hasan (Khulna Univ. of Engr. and Tech.)
BanglaDocAtlas: A Multi-Class Annotated Dataset for Complex Bangla Document Layout Analysis
Md Safayat Hossain, Jannatul Ferdous (United Intl. Univ.), Md Raihan Uddin (Wichita State Univ.), K. M. A. Hossain, M. I. Ahmed, M. A. Rahman (United Intl. Univ.), Asif Sushmit (Bengali.AI), Farig Sadeque, Swakkhar Shatabda (BRAC University), Abu Asaduzzaman (Wichita State Univ.)
 

4-5: GenAI Opportunities & AI Challenges Session (17:30-19:30)

Organizers: V. Gadepally, D. Burrill & C. Prothmann  
MI300A vs H100, LLM Concerns for Applications at NRL, and MCP Testing
William “Connor” Horne (NRL)
Trace Replay Simulation of MIT SuperCloud Dataset for Studying Optimal Policies for Sustainability [Outstanding Short Paper Award]
Wesley Brewer, Matthias Maiterth (ORNL), Damien Fay (HPE)
Introduction: Advancing AI Challenges for the United States Department of the Air Force
Christian Prothmann, Vijay Gadepally, Jeremy Kepner, Koley Borchard, Luca Carlone, Zachary Folcik,J. Daniel Griffith, Michael Houle, Jonathan P. How, Nathan Hughes, Ifueko Igbinedion, Hayden Jananthan, Tejas Jayashankar, Michael Jones, Sertac Karaman, Binoy G. Kurien, Alejandro Lancho, Giovanni Lavezzi, Gary C. F. Lee, Charles E. Leiserson, Richard Linares, Lindsey McEvoy, Peter Michaleas, Chasen Milner, Alex Pentland, Yury Polyanskiy, Jovan Popovich, Jeffrey Price, Tim W. Reid, Stephanie Riley, Siddharth Samsi, Peter Saunders, Olga Simek, Mark S. Veillette, Amir Weiss, Gregory W. Wornell, Daniela Rus, Scott T. Ruppel (MIT)
Data-Driven Radio-Frequency Signal Separation Challenge
Yury Polyanskiy (MIT)
Tornado Network Challenge
Mark Veillette (MIT Lincoln Laboratory), Peter Saunders (USAF)
AI Innovation in Space Challenge
Giovanni Lavezzi, Richard Linares (MIT)
Evaluating Long-Context LLM Architectures with Controlled Synthetic Benchmarks Challenge
Olga Simek (MIT Lincoln Laboratory)
Predicting LLM Inference Server Request Capacity
Daniel J. Burrill, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alexander Bonn, Chansup Byun, Michael Houle, Matthew Hubbell, Michael Jones, Piotr Luszczek, Peter Michaleas, Guillermo Morales, Julie Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee (MIT Lincoln Laboratory), Vijay Gadepally (MIT)
   

Friday, September 19

 

5-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther  
Keynote Talk: Performance Engineering with MPI
Bill Gropp (NCSA)
 

5-1: Mixed Precision Session (11:00-12:15)

Co-Chairs: P. Luszczek & C. Byun  
Invited Talk: Is Mixed Precision Computing really the Top Priority?
Hartwig Anzt (TU Munich)
Invited Talk: Reducing Numerical Precision Requirements in Quantum Chemistry Calculations
William Dawson (RIKEN)
Invited Talk: Why Exact Dot Products Obviate the Need for Mixed Precision
John Gustafson (ASU)
Invited Talk: Mixed Feelings about Mixed Precision
Hatem Ltaief (KAUST)
Invited Talk: Mixed Precision or Mixed Storage? User-Guided Compiler Transformations to Change the Data Layout On-the-Fly
Tobias Weinzierl (Durham Univ.)
 

5-2: Mixed Precision Session (12:30-13:45)

Co-Chairs: P. Luszczek & R. Muri  
Invited Talk: Emulation Method for Matrix Multiplication
Katsuhisa Ozaki (Shibaura Inst. of Tech.)
A variable-precision implementation of the ADER-DG algorithm
Marc Marot-Lassauzaie, Michael Bader (Tech. Univ. Munich)
GPU-Accelerated, Mixed Precision GMRES(m) with Varied Restarts [Outstanding Student Paper Award]
Abir Haque, Suzanne Shontz, Xuemin Tu (Univ. of Kansas)
Performance and Numerical Aspects of Decompositional Factorizations with FP64 Floating-Point Emulation in INT8
Piotr Luszczek, Vijay Gadepally, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel J. Burrill, Chansup Byun, Michael Houle, Matthew Hubbell, Michael Jones, Peter Michaleas, Guillermo Morales, Julia Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Jeremy Kepner (MIT Lincoln Laboratory)
Adaptive Spectral Block Floating Point for Discontinuous Galerkin Methods
Shivam Sundriyal, Markus Büttner (Univ. of Bayreuth), Christoph Alt, Tobias Kenter (Paderborn Univ.), Vadym Aizinger (Univ. of Bayreuth)
 

5-P2 (13:45-14:45): Embedded & GPU Poster Session

Chair(s)/Host(s): S. Shankar  
Decision Support for Sustainable Agriculture Using I2C Sensors
Anita Esmaeilian, Kishwar Ahmed (Univ. of Toledo)
ORB Feature Extraction on Embedded Platforms: A Heterogeneous CPU-GPU-PVA Approach
Hamid Moghadaspour (Univ. of Coimbra), Nuno Neves (Univ. of Lisbon), Oscar Ferraz, Gabriel Falcao (Univ. of Coimbra)
Accelerating Temporal Triangle Counting and Betweenness Centrality on GPUs
Tuteja Trimansingh Parvindersingh, Venkata Kalyan Tavva (IIT Ropar), Subhasis Banerjee, Chiranjib Sur (Shell India Markets Pvt. Ltd.)
SCoDa: Scalable Community Detection in Data Streams
Akanksha Dwivedi, Prashant Srivastav, Dip Sankar Banerjee (IIT Jodhpur)
CTAM Tool for Hyperscaler Qualification
Tommy Yan, Rajat Madhusudan, Vani Pulendra, Anna Mary Mathew (Silicon Cloud HIE)
 

5-3: Cyber Analysis and Secure Computing Session (14:15-15:30)

Co-Chairs: D. Cousins & R. Vuduc  
Invited Talk: Private Database Analytics with PAC Privacy
Srini Devadas (MIT)
Accelerating Multi-Party Computation Using Heterogeneous Systems [Outstanding Student Paper Award]
Xiteng Yao, Shining Yang, Mayank Varia, Martin Herbordt (Boston Univ.)
Optimizing Local Computation in Secure Matrix Multiplication for Outsourced Neural Networks [Outstanding Paper Award]
J. Parker Diamond, Andrea Lin, R. Nicholas Cunningham (MIT Lincoln Laboratory), Soamar Homsi (AFRL), John Darby Mitchell, Aseemit Pandey, Emily Shen (MIT Lincoln Laboratory)
A Framework For The Iterative Solution of Sparse Linear Systems on Hybrid Architectures Using Homomorphic Encryption
Lior Horesh, Vasileios Kalantzis, Barry M. Trager, Shashanka Ubaru (IBM Research)
Secure Virtual Network Embedding Through Fully Homomorphic Encryption
David Bruce Cousins, Carlo Pascoe (Duality Tech.), Erik Kline (USC)
 

5-4: Embedded Computing Session (15:45-17:15)

Co-Chairs: D. Cousins & E. Schnetzer  
Invited Talk: AI and Biodiversity
Sara Beery (MIT)
On the Adaptation of Mixed-Radix Fast Fourier Transform for Resource-Constrained Environments [Outstanding Student Paper Award]
Atharva Gondhalekar, Paul Sathre, Wu-chun Feng (Virginia Tech)
Comparative Analysis of RISC-V Softcore and Hardcore Processors for Space Computing [Outstanding Student Paper Award]
Ni Nyoman Dhinar Gayatri, Alan D. George (Univ. of Pittsburgh)
TRACER Software Switching at Waveform Timescales
Connor Imes, Dong In D. Kang, Matthew French, John Paul Walters (USC Information Sciences Institute)
Hardware-Accelerated Transformer Framework for Real-Time Battery SoH Estimation
Talha Coskun (Univ. of Illinois Urbana-Champaign), Hiruna Vishwamith (Univ. of Moratuwa), Murat Isik (Stanford Univ.), I. Can Dikmen (Istinye Univ.)
A Qualifiable GPU Sharing Approach for AI Workloads in Critical Systems
Marc Solé i Bonet, Jannis Wolf (Barcelona Supercomputing Ctr.), Aridane Álvarez Suárez (fentISS), Leonidas Kosmidis (Barcelona Supercomputing Ctr.)
 

5-5: AI for Performance Engineering Session (17:30-19:30)

Organizers: H. Nguyen & D. Burrill  
UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC [Outstanding Student Paper Award]
Tomer Bitan (Technion), Tal Kadosh (Ben-Gurion Univ. / IAEC), Erel Kaplan, Shira Meiri (Technion), Le Chen (Argonne NL), Peter Morales, Niranjan Hasabnis (Code Metal), Gal Oren (Stanford Univ.)
Web-Based Intelligent Decision Support System for Real-Time Toll Plaza Management and AI-Driven Operational Optimization
Pattarapon Klaykul, Wilaiporn Lee, Kanabadee Srisomboon, Luepol Pipanmekaporn, Akara Prayote (King Mongkut’s Univ.)
CRAMP: Categorizing Classifiers and Regressors for Scalable Parallelism on Distributed and Multicore Systems
Baidya Nath Saha, Pavan Sarvaiya, Wali Mohammad Abdullah, Md. Morshedul Islam (Concordia Univ. of Edmonton)
RAILS: Retrieval-Augmented Intelligence for Learning Software Development
Wali Mohammad Abdullah, Md. Morshedul Islam, Devraj Parmar, Happy Hasmukhbhai Patel, Sindhuja Prabhakaran, Baidya Saha (Concordia Univ. of Edmonton)
Towards Automated Reasoning Chains for Verification of LLM-Generated Scientific Code
Quentin Oschatz, Naifeng Zhang (Carnegie Mellon Univ.), Mike Franusich (SpiralGen), Franz Franchetti (Carnegie Mellon Univ.)
P4OMP: Retrieval-Augmented Prompting for OpenMP Parallelism in Serial Code
Wali Mohammad Abdullah (Concordia Univ. of Edmonton), Azmain Kabir (Univ. of Manitoba)
Towards -OmL: A Deep Learning Based Approach to Outperform Compiler Defaults
Hafsah Shahzad (Boston Univ.), Ahmed Sanaullah, Sanjay Arora, Ulrich Drepper (Red Hat), Martin Herbordt (Boston Univ.)
   

IEEE HPEC 2025