30th Annual
IEEE High Performance Extreme Computing Virtual Conference
14 - 18 September 2026

HPEC 2026 Agenda

To be announced!

Below is the agenda from HPEC 2025.   Final papers published to the HPEC website can be found within the agenda below where there is a link associated with the paper title.  The papers that were chosen to be published in the IEEE Xplore Digital Library can be found here: 2025 IEEE High Performance Extreme Computing Conference (HPEC) – Conference Table of Contents | IEEE Xplore.

2025 AGENDA

Speaker/Presenting Author in Italics

DayMondayTuesdayWednesdayThursdayFriday
10:30-11:00amSession 1-K: KeynoteSession 2-K: KeynoteSession 3-K: KeynoteSession 4-K: KeynoteSession 5-K: Keynote
11:00am-12:15pmSession 1-1: Advanced Processor Architectures Session 2-1: Bridging Quantum and HPC Session 3-1: Scaling Research Computing Education  Session 4-1: High Performance Computing Session 5-1: Mixed Precision
12:15-12:30pmBreakPoster Session 1-P1 (12:15-13:15): New Application FrontiersBreakPoster Session 2-P1 (12:15-13:15): Graph AI & SparseBreak Tutorial Session 3-T (12:15-15:45): Spiral TutorialBreakPoster Session 4-P1 (12:15-13:15): AI/ML/GenAIBreak 
12:30-1:45pmSession 1-2: Neuromorphic AISession 2-2: Case Studies, Benchmarking, and ToolsSession 3-2: AI in the WorldSession 4-2: High Performance ComputingSession 5-2: Mixed Precision
1:45-2:15pmBreakPoster Session 1-P2 (13:45-14:45): High Performance Data AnalysisBreak BreakPoster Session 3-P2 (13:45-14:45): High Performance ComputingBreakPoster Session 4-P2 (13:45-14:45): AI/ML/GenAIBreakPoster Session 5-P2 (13:45-14:45): Embedded & GPU
2:15-3:30pmSession 1-3: ASIC and FPGA AdvancesSession 2-3: Quantum and Non-Deterministic ComputingSession 3-3: Graph AI & SparseSession 4-3: High Performance ComputingSession 5-3: Cyber Analysis and Secure Computing
3:30-3:45pmBreakBreakBreakBreakBreak
3:45-5:00pmSession 1-4: FastCode Session 2-4: Graph AI & Sparse Session 3-4: Graph AI and Sparse Session 4-4: AI/ML/GenAI Session 5-4: Embedded Computing
5:00-5:30pmBreakBreakBreakBreakBreak
5:30-7:30pmSession 1-5: BRAINS – Building Resilience through Artificial Intelligence for Networked SystemsSession 2-5: GraphBLAS BoFSession 3-5: Graph ChallengeSession 4-5: GenAI Opportunities & AI ChallengesSession 5-5: AI for Performance Engineering

Monday, September 15

1-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther

Keynote Talk: Enabling Advances in OceanAI

Nick Rotker (MITRE)

1-1: Advanced Processor Architectures Session (11:00-12:15)

Co-Chairs: P. Luszczek & C. Byun

Invited Talk:

Michael Foertsch (Q.ANT)
Encoded Time-Series Model Training with UNet Running on Wafer Scale Engine

Vyacheslav Romanov (NETL)
Lincoln AI Computing Survey (LAICS) and Trends

Albert Reuther, Peter Michaleas, Michael Jones, Vijay Gadepally, Jeremy Kepner (MIT Lincoln Laboratory Supercomputing Center)
System-Level Performance Modeling of Photonic In-Memory Computing

Jebacyril Arockiaraj, Sasindu Wijeratne (USC), Sugeet Sunder (USC Information Sciences Institute), Md Abdullah-Al Kaiser, Akhilesh Jaiswal (Univ. of Wisconsin), Ajey P Jacob (USC Information Sciences Institute), Viktor Prasanna (USC)

1-P1 (12:15-13:15): New Application Frontiers Poster Session

Chair(s)/Host(s): K. Keville

Multi-Stage Stochastic Programming for Heavy-Duty Electric Truck Routing Under Public Charging Congestion Uncertainty

Ziyan Li, Nikolay Aristov, Antoine Germain, Elenna R. Dugundji (MIT)

Enabling Heterogeneous Performance Analysis for Scientific Workloads [Outstanding Short Paper Award]  

 

Maksymilian Graczyk (CERN), Vincent Desbiolles (HES-SO), Stefan Roiser (CERN), Andrea Guerrieri (HES-SO and EPFL)
Weighted Histogram Matching for Improved Automated Synthetic Aperture Radar (SAR)-optical Image Registration

Kerri Prinos, Aditi Mungale, Dr. Karen Gettings, Dr. Tair Akhmejanov, Devanshu Mehta (MIT Lincoln Laboratory)
Towards a Framework for Etymological Analysis of Literary and AI Works

Mike Teodorescu (Univ. of Washington), Cecilia Speranța Bolea (IIT), Horia Teodorescu (Technical Univ. Iasi)

An Algorithmically Determined Exchange-Correlation (ADEXC) Method for Electronic Structure Calculations  

 

Ivan Williams, Eric Polizzi (UMass Amherst)

Scaling Performance of Large Language Model Pretraining [Best Short Paper Award]  

Alexander Interrante-Grant, Carla Varela-Rosa, Suhaas Narayan, Chris Connelly, Albert Reuther (MIT Lincoln Laboratory)

Metadata Guided Pose Estimation for 3D Reconstruction  

 

William Farthing (Emory Univ.), Compton Ross (Univ. of Mississippi), Jaired Collins, Jing-Ru C. Cheng (U.S. Army Corp of Engineers)

Neuromorphic Processor Employing FPGA Technology with universal interconnections  

 

Pracheta Harlikar, Abdel-Hameed Badawy (New Mexico State Univ.), Prasanna Date (ORNL)

1-2: Neuromorphic AI Session (12:30-13:45)

Co-Chairs: M. Barnell & H. Nguyen

Artificial Intelligence Performance and Radiation Effects in Neuromorphic NorthPole Hardware

Victor M. Vergara (BlueHalo-AFRL), Francisco O. Viramontes, Amanda E. Romero (COSMIAC Research Center), Matthew E. Spear, Evan T. Kain, Windy S. Slater, Heather M. Quinn, Qing Wu (AFRL)
Physics-Informed Neural Networks for Low-Power Real-Time Edge Biosensing Applications

Soheli Farhana (Harvard Univ.)
Analysis and Optimization of Spiking Neural Network Simulations on GPUs

Doğu Kocatepe, Işıl Öz (Izmir Institute of Technology)
Exploring Neuromorphic Computing with Loihi-2 for High-Performance CFD Simulations

Talha Coskun (Univ. of Illinois Urbana-Champaign), Hiruna Vishwamith (Univ. of Moratuwa), Murat Isik (Stanford Univ.), I. Can Dikmen (Istinye Univ.)
Neural-Inspired Enhancing Spiking Graph Convolutional Networks

Fernando Vera Buschmann, Horacio Roststein, Vincent Oria (New Jersey Inst. of Tech.)

1-P2 (13:45-14:45): High Performance Data Analysis Poster Session

Chair(s)/Host(s): K. Keville

Accelerating Push-Relabel Algorithm on GPU via Two-Level Parallelism Paradigm and Efficient CSR Designs

Chou Ying Hsieh, Po-Chieh Lin, Sy-Yen Kuo (National Taiwan Univ.)
Improving Statistical Characterization of Data Tensors with the Generalized Canonical Polyadic Tensor Decomposition

Matthew Merris, Tim Andersen (Boise State Univ.)

Comparative Analysis of Classical and Deep Learning Features for Texture Image Classification  

 

Murad Hossen (Univ. of Houston)

Aspects of HPC within Conformal Recommendation Systems: Ideas and Challenges  

 

Stanisław M. S. Halkiewicz (AGH Univ. of Cracow), Maciej Kuczyński (Czestochowa Univ. of Tech.)
Enhancing the Real-Time Solutions of Parametric Linear Systems on a GPU through Hybrid Coarse-Grained Transprecision Computing

Hamid Noori, Hans Vandierendonck, Roger Woods (Queen’s Univ. Belfast), Nick Polydorides (Univ. of Edinburgh)

AAPA: An Archetype-Aware Predictive Autoscaler with Uncertainty Quantification for Serverless Workloads on Kubernetes  

 

Guilin Zhang (George Washington Univ.), Srinivas Vippagunta, Raghavendra Nandagopal, Suchitra Raman, Jeff Xu, Marcus Pfeiffer, Shreeshankar Chatterjee (Workday), Ziqi Tan, Wulan Guo (George Washington Univ.), Hailong Jiang (Youngstown State Univ.)
Stable Iterative Solvers for Ill-Conditioned Linear Systems and Least Squares

Vasileios Kalantzis, Mark S. Squillante, Chai Wah Wu (IBM Research)

1-3: ASIC and FPGA Advances Session (14:15-15:30)

Co-Chairs: L. Zaidenberg & K. Thurmer

NTT-SAA: Exploring NTT Acceleration with 2-D Systolic Array Architecture on FPGAs

Ashwajit Singh (IIT Bombay), Zhihan Xu, Viktor K. Prasanna (USC)
Evaluating AMD-Xilinx Frameworks for Deep-Learning Acceleration on Versal [Outstanding Student Paper Award]

Peter Drum, Alan D. George (NSF SHREC)
Fast FPGA-Based Implementation of the QP-Dyn Stream Cipher using High-Level Synthesis

Paolo Palazzari (ENEA), Luigi Accardi (Volterra Univ.), Antonio Mastrandrea, Pasquale Tommasino, Alessandro Trifiletti (Sapienza Univ.)
A Scalable Code Generation Flow for Heterogeneous Parallel RTL Simulation using MLIR

Jie Tong, Zhengxiong Li, Umit Yusuf Ogras, Tsung-Wei Huang (Univ. of Wisconsin)
Accelerating Dynamic Image Graph Construction on FPGA for Vision GNNs

Anvitha Ramachandran, Dhruv Parikh, Viktor Prasanna (USC)

1-4: FastCode Session (15:45-17:15)

Co-Chairs: Bruce Hoppe

Invited Talk: Fastcode: An Open-Source Community for Making Software Performance Engineering Easy and Fun

Bruce Hoppe (MIT)
Invited Talk: Taskflow — A General-Purpose Task-Parallel Programming System

Tsung-Wei Huang (Univ. of Wisconsin)
Invited Talk: OpenCilk — A Modular and Extensible Software Infrastructure for Fast Task-Parallel Code

I-Ting Angelina Lee (WUSTL)
Invited Talk: Multithreaded Parallel Python Through OpenMP Support in Numba

Tim Mattson (Human Learning Group)

1-5: BRAINS – Building Resilience through Artificial Intelligence for Networked Systems Session (17:30-19:30)

Organizers: T. Hardjono & S. Pisharody

Invited Talk: Developing Trust in the Supply Chain

Guy Fedorkow (Juniper/HPE)
Invited Talk: Using AI to Achieve Peak Supply Chain Health

Karyl Fowler (Tradeverifyd)
Invited Talk: Supply Chain Transparency as Entailment for Device Attestation

Ned Smith (Intel)
Towards an Algorithm-based Approach for Soft Error Tolerance using Interval Arithmetic [Best Paper Award]

Larry Tang, Varun Kumar, Matt Ngaw, Siddharth Singh, Devdutt Nadkarni, Lohith Tummala, Ken Mai, Franz Franchetti (Carnegie Mellon Univ.)
Detecting Expert-Written Comments in Stack Exchange

Himani Musku (Carnegie Mellon Univ.), Alea Ritchie (Stanford Univ.), Nour Jedidi, Rohan Leekha, Courtland VanDam (MIT Lincoln Laboratory)
Strategic Cyber Defense via RL-Guided Combinatorial Auctions

Mai Pham, Vikrant Vaze, Peter Chin (Dartmouth Coll.)
AI-Powered Network Energy Optimization Machine Learning Approaches to Reducing Network Power Consumption

Het Mehta (Cisco Systems), Sambu Patach Arrojula (Samsung)

Tuesday, September 16

2-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther

Keynote Talk: Effective Human-Machine Partnerships in High Stakes Settings

Julie Shah (MIT)

2-1: Bridging Quantum and HPC Session (11:00-12:15)

Co-Chairs: Devesh Tiwari

Invited Talk: Tightening the Integration of Quantum Computers with HPC Systems

Travis Humble (ORNL)
Invited Talk: Enabling Quantum Utility through a System Toolkit

Tirthak Patel (Rice Univ.)
Invited Talk: CUDA-Q: Defining a Tightly-Integrated System Software Stack for Quantum-Accelerated HPC

Alex McCaskey (NVIDIA)

2-P1 (12:15-13:15): Graph AI & Sparse Poster Session

Chair(s)/Host(s): K. Cain

Cyber Orbits of Large Scale Network Traffic

Jeremy Kepner, Hayden Jananthan, Chasen Milner, Michael Houle, Michael Jones, Peter Michaleas, Alex Pentland (MIT)
Accelerating AI Development with Cyber Arenas

William Cashman, Chasen Milner (USAF), Michael Houle, Michael Jones, Hayden Jananthan, Jeremy Kepner, Peter Michaleas, Alex Pentland (MIT)
Optimizing Sparse Matrix-Vector Multiplication on GPUs using the Mathematics of Arrays

Stephen Thomas (Lehigh Univ.), Lenore Mullin (Univ. of Albany)

Sampling to Scale: Performance Trade-offs in Approximate Triangle and Square Counting  

 

Shubhashish Kar, Shaikh Arifuzzaman (UNLV)

Spectral Sparsification of Edges for Efficient Clique Counting  

 

Aaron Schindler (Univ. of Utah), Hari Sundar (Tufts Univ.)
Degree Matrix Comparison for Graph Alignment

Ashley Wang, Peter Chin (Dartmouth Coll.)

2-2: Case Studies, Benchmarking, and Tools Session (12:30-13:45)

Co-Chairs: X. Sun & J. Ghanem

Invited Talk: Rigorous Mathematics in Data-driven Methods for Safe Autonomy

Chuchu Fan (MIT)
AOCL-CST – A new CPU Stress Test library for AMD CPUs

S. Biplab Raut (AMD)
Performance–Energy Characterization of ML Inference on Heterogeneous Edge AI Platforms

Palash Kohli (BITS Pilani), Rakshith Jayanth, Neelesh Gupta, Haoyang Fan, Viktor Prasanna (USC)
A Fully Adaptive Radau Method for the Efficient Solution of Stiff Ordinary Differential Equations at Low Tolerances

Shreyas Ekanathan (Lexington High School), Oscar Smith (JuliaHub), Christopher Rackauckas (MIT)
Evaluation of Habitat Robotics using Large Language Models

William Li, Lei Hamilton, Kaise Al-natour, Sanjeev Mohindra (MIT Lincoln Laboratory)

2-3: Quantum and Non-Deterministic Computing Session (14:15-15:30)

Co-Chairs: B. Sroka & I. DeTore

SYCL for Performance Portability: Application Experience with Coupled Cluster Formalism in Quantum Chemistry on Exascale Systems [Outstanding Paper Award]

Abhishek Bagusetty (Argonne National Laboratory), Ajay Panyala (PNNL), Alvaro Vazquez Mayagoitia (Argonne National Laboratory), John K. Holmen (ORNL), Kevin Harms (Argonne National Laboratory)
Implementation of Tensor Network Simulation TN-Sim under NWQ-Sim

Aaron C. Hoyt, Jonathan S. Bersson, Sean Garner (Univ. of Washington), Chenxu Liu, Ang Li (PNNL)
Partition-based Surface Code Compilation

Hanjing Xu (Purdue Univ.), Xiaoyuan Liu, Ankit Kulshrestha, Hayato Ushijima-Mwesigwa (Fujitsu Research of America)
QAOA Parameter Transferability for Maximum Independent Set using Graph Attention Networks

Hanjing Xu (Purdue Univ.), Xiaoyuan Liu (Fujitsu Research of America), Alex Pothen (Purdue Univ.), Ilya Safro (Univ. of Delaware)
A Hybrid Classical-Quantum Model for QSAR-Based Biodegradability Prediction

Batuhan Hangun, Oguz Altun (Yıldız Technical Univ.), Onder Eyecioglu (Bolu Abant İzzet Baysal Univ.)

2-4: Graph AI & Sparse Session (15:45-17:15)

Co-Chairs: X. Sun & J. Ghanem

Scalable Graph Algorithms on Distributed UpDown Accelerators

Brian Wheatman, Andrew A. Chien (Univ. of Chicago)
2D Distributed Label Propagation on 400 GPUs

George M. Slota, Michael Mandulak, Ujwal Pandey, Anthony Fabius (RPI)
Accelerating Sparse Deep Learning via Multi-Layer Tensor Reordering and Partitioning

Gunduz Vehbi Demirci (Imagination Tech.), Cagatay Dikici (Wayve), Tim Atherton (Imagination Tech.)
Julia GraphBLAS with Nonblocking Execution [Outstanding Paper Award]

Pascal Costanza (Independent Researcher), Timothy G. Mattson (Univ. of Bristol), Raye Kimmerer (NERSC LBNL), Benjamin Brock (Intel)
A parallel push-relabel maximum flow algorithm in LAGraph and GraphBLAS

Darin A. Peries, Timothy A. Davis (Texas A&M Univ.)
GraphBLAS Mathematical Opportunities: Parallel Hypersparse, Matrix Based Graph Streaming, and Complex-Index Matrices

Hayden Jananthan, Jeremy Kepner, Michael Jones, Vijay Gadepally, Michael Houle, Peter Michaleas, Chasen Milner, Alex Pentland (MIT)

2-5: GraphBLAS BoF Session (17:30-19:30)

Organizers: T. Mattson, B. Brock & S. McMillan

GraphBLAS APIs: The Math Spec

Tim Mattson (Human Learning Group)
Keynote – Postgres and GraphBLAS

Michel Pelletier (OneSparse)
SuiteSparse GraphBLAS and LAGraph: Progress Report and Future Direction

Tim Davis (Texas A&M)

Wednesday, September 17

3-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther

Keynote Talk: Building AI That Users Trust

Ashley Conard (Microsoft)

3-1: Scaling Research Computing Education Session (11:00-12:15)

Co-Chairs: J. Mullen, L. Milechin & H. Jananthan

Invited Talk: Research/Advanced Computing Roles: Skillfully Facing the Challenges

Robert M. Freeman, Jr. (Harvard University & CaRCC)
Invited Talk: Skill Inventories, What They Are and Why We Need Them

Weronika Filinger (Edinburgh Parallel Computing Center)
Invited Talk: Cataloguing the Training Material Landscape: Skills, Gaps, and Resources

Jeremy Cohen (Imperial Col.)
Invited Talk: Designing an Integrated Ecosystem to Develop Research Computing and Digital Technical Professionals

Julia Mullen (MIT LL)

3-2: AI in the World Session (12:30-13:45)

Co-Chairs: K. Gettings & M. Barnell

Invited Talk: Learning-Guided Optimization for Mobility

Cathy Wu (MIT)
Sustainably Modeling a Sustainable Future Climate

Rabab Alomairy (MIT), Sameh Abdulah (KAUST), Qinglei Cao (St. Louis Univ.), Marc G. Genton, David E. Keyes, Hatem Ltaief (KAUST)
Predicting Ports Congestion by Utilizing State Space Models

Nikolay Aristov, Elenna R. Dugundji (MIT-CTL)
Scaling Regime-Aware Forecasting: Distributed Shifting Seasonal Matrix Factorization

Jacob Munson, Breschine Cummins (Montana State Univ.)
AIMS: An Adaptive Intelligent Multi-Objective Scheduler Powered by Digital Twins

Kyrian Adimora, Hongyang Sun (Univ. of Kansas)

3-P2 (13:45-14:45): High Performance Computing Poster Session

Chair(s)/Host(s): P. Luszczek

When Structure is Silent: Opportunities for Algorithmic Dispatch in Linear Algebra

Emmanuel Lujan, Alan Edelman (MIT)
Performance Modeling of Heterogeneous Edge-Cloud Systems with Machine Learning

Md Raihan Uddin, Abu Asaduzzaman, Sonu Gangadhar Gowda (Wichita State Univ.)
An FFT-based Preconditioner for Conjugate Gradient Pressure Solvers in Complex Domains

Xin Kai Lee, Gregory LeClaire Wagner, Simone Silvestri, Raffaele Ferrari (MIT)
A Scalable Quantum Dynamical Approach for Calculating Collisional Molecular Properties

Prajwal Niraula, Laurent Wiesenfeld, Julien de Wit (MIT), Iouli Gordon, Robert Hargreaves (Harvard Univ.), Jeremy Kepner, Deborah Woods, Cooper Loughlin (MIT Lincoln Laboratory)
Scalable Bayesian Nonparametric Ensemble (BNE) for Spatio-temporal Air Pollution Predictions using High-Performance Computing

Vijay Kumar, Jaime Benavides (Brown University), Carlos Carrillo-Gallegos (Columbia Univ.), Gil Speyer (Arizona State Univ.), Marianthi-Anna Kioumourtzoglou (Brown University)
Investigating the Impact of Algorithms and Hardware on Machine Learning Models in HPC Systems

Christian C. Thompson, Abu Asaduzzaman, Md Raihan Uddin (Wichita State Univ.)
SARComp: High-Performance Algorithms for Onboard SAR from FFT Kernels to Matched Filtering

Maron Schlemon (German Aerospace Center), Martin Schulz (Tech. Univ. Munich), Rolf Scheiber (German Aerospace Center)

Streamed Multi-Format Sparse Matrix Vector Multiplication for FPGA  

 

Spencer Smith, Richard Veras (Univ. of Oklahoma)

3-3: Graph AI & Sparse Session (14:15-15:30)

Co-Chairs: N. Pitsianis & C. Byun

pdGRASS: A Fast Parallel Density-Aware Algorithm for Graph Spectral Sparsification [Best Student Paper Award]

Tiancheng Zhao (Georgia Inst. of Tech.), Zekun Yin, Huihai An (Shandong Univ.), Xiaoyu Yang (China Univ. of Petroleum-Beijing), Zhou Jin (Zhejiang Univ.), Jiasi Shen (HKUST), Helen Xu (Georgia Inst. of Tech.)
Differentiable Graph Centrality

Georgios Kollias, Vassilis Kalantzis (IBM Research)
Performance Analysis of the Parallel Shared-Memory Sparse Matrix-Vector Multiplication on Unstructured Matrices

Kobe Bergmans, Karl Meerbergen, Raf Vandebril (KU Leuven)
HiPerMotif: Novel Parallel Subgraph Isomorphism in Large-Scale Property Graphs

Mohammad Dindoost, Oliver Alvarado Rodriguez, Bartosz Bryg, Ioannis Koutis, David A. Bader (New Jersey Inst. of Tech.)
Characterization of Sparsity-aware Parallelization of Jaccard Similarity in Graph Datasets

Atharva Gondhalekar, Paul Sathre, Wu-chun Feng (Virginia Tech)

3-4: Graph AI and Sparse Session (15:45-17:15)

Co-Chairs: N. Pitsianis & C. Byun

Enhancing Graph Partitioning with Reinforcement Learning-based Initialization

Chedi Morchdi (Texas A&M Univ.), Cheng-Hsiang Chiu, Wan Luan Lee, Tsung-Wei Huang (Univ. of Wisconsin), Yi Zhou (Texas A&M Univ.)
Benchmarking Deep Learning with Representative ONNX Subgraphs

Marika E. Schubert (Univ. of Pittsburgh), David Langerman (NSF SHREC), Evan W. Gretok, Ian Peitzsch, Calvin B. Gealy, Jefferson Boothe, Alan D. George (Univ. of Pittsburgh)
A Locality Sensitive Hashing Based Algorithm to Accelerate Neighborhood Search in Graph Neural Operators

Mariam Hassan, Sanmukh Kuppannagari (Case Western Reserve Univ.)
Power Iteration with Probabilistic Updates for Systems with Heterogeneous Performance

Soumyadip Ghosh, Lior Horesh, Vasileios Kalantzis, Georgios Kollias, Yingdong Lu, Tomasz Nowicki, Shashanka Ubaru (IBM Research)
Designing Parallel Algorithms for Community Detection using Arachne

Fuhuan Li, Zhihui Du, David Bader (New Jersey Inst. of Tech.)
GCN-Driven CUDA Parameter Optimization for Parallel Triangle Counting in Graphs

Hasan Serdar Arikan, Rakibul Hassan, Shubhashish Kar (Univ. of Nevada Las Vegas), Doru Thom Popovici (LBNL), Shaikh Arifuzzaman (Univ. of Nevada Las Vegas)

3-5: Graph Challenge Session (17:30-19:30)

Organizers: J. Kepner & A. Reuther

Combining Performance and Productivity: Accelerating the Network Sensing Graph Challenge with GPUs and Commodity Data Science Software

Siddharth Samsi, Dan Campbell, Emanuel Scoullos, Oded Green (NVIDIA)
SANST: Sensing Anonymized Network via Sorted Triplets

Jianyu Wang, Wenzi Tang, Chenglong Shi, Zhe Zhang (Guangxi Univ.), Dan Chen (National Univ. Singapore), Miaojiang Chen, Wenjing Xiao (Guangxi Univ.)
Anonymized Network Sensing using C++26 std::execution on GPUs

Michael Mandulak (RPI), Sayan Ghosh, S M Ferdous, Mahantesh Halappanavar (PNNL), George Slota (RPI)
Geans: A GPU-accelerated Framework for Efficient End-to-End Anonymized Network Sensing

Jun Mai, Qinggang Wang, Yu Huang, Pengcheng Yao, Long Zheng, Xiaofei Liao, Hai Jin (Huazhong Univ. of Science and Tech.)
DBOS Network Sensing: A Web Services Approach to Collaborative Awareness

Sophia Lockton, Jeremy Kepner, Michael Stonebraker, Hayden Jananthan, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel Burrill, Chansup Byun (MIT), Timothy Davis (Texas A&M), Vijay Gadepally, Michael Houle, Matthew Hubbell, Michael Jones, Piotr Luszczek, Peter Michaleas, Lauren Milechin, Chasen Milner, Guillermo Morales, Julie Mullen (MIT), Michel Pelletier (OneSparse), Alex Poliakov (DBOS), Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Alex Pentland (MIT)
Interactive Trillion Packet Anonymized Network Analysis with the GraphBLAS

Chasen Milner (USAF), Michael Houle, Hayden Jananthan, Michael Jones, Jeremy Kepner, Peter Michaleas, Inna Voloshchuk, Alex Pentland (MIT)
Towards Efficient Sparse Deep Neural Network Inference via Multi-level Concurrency Orchestration

Ming Dun, Jie Zhou (Inst. of Computing Tech, CAS), Huawei Cao (UCAS), Shuhan Song, Yiming Sun, Mingyu Yan, Xiaochun Ye (Inst. of Computing Tech, CAS)
Scaling Triangle Counting and K-Truss on the UpDown Architecture

Jiya Su, Alexander Fell, Andronicus Rajasukumar (Univ. of Chicago), David F. Gleich (Purdue Univ.), Andrew A. Chien (Univ. of Chicago)
PRISM: Practical In-Memory Acceleration for Subgraph Matching at Scale

Deting Chen, Yu Huang, Yi Huang, Binbin Lin, Yi Zhang, Long Zheng, Xiaofei Liao, Hai Jin (Huazhong Univ. of Science and Tech.)
BUG: Balanced DFS-Based Subgraph Matching with a Reuse Strategy on GPUs

Zicang Xu, Lei Zou (Peking Univ.)
InfraredGP: Efficient Graph Partitioning via Spectral Graph Neural Networks with Negative Corrections

Meng Qin (Pengcheng Laboratory), Weihua Li (Beihang Univ.), Jinqiang Cui (Pengcheng Laboratory), Sen Pei (Columbia Univ.)

Thursday, September 18

4-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther

Keynote Talk: SPACE MICE: The Next Generation Data Systems

Joshua Patterson (NVIDIA)

4-1: High Performance Computing Session (11:00-12:15)

Co-Chairs: D. Cousins & H. Sadasivan

AGCRS: An Adaptive Generalized Storage Scheme for Large Sparse Tensors

Md Mehrab Hossain Opi, K. M. Azharul Hasan (Khulna Univ. of Engr. and Tech.)
Accelerating Supercomputing: AI-Hardware-Driven Innovation for Speed and Efficiency

Jack Dongarra (Univ. of Tennessee), John Gunnels, Harun Bayraktar, Azzam Haidar, Dan Ernst (NVIDIA)
On the Landscape of Scientific Computing Libraries in Python

Niteya Shah (Virginia Tech), Pi-Yeuh Chuang (Argonne National Laboratory), Paul Sathre, Wu-chun Feng (Virginia Tech)
Load Imbalance in HPC Applications: Improved Profiling and New Ways to Use Wasted Cycles

Shining Yang, Xiteng Yao (Boston Univ.), Grace Nansamba, Amr Akmal Abouelmagd, Anthony Skjellum (Tennessee Tech), Martin Herbordt (Boston Univ.)

4-P1 (12:15-13:15): AI/ML/GenAI Poster Session

Chair(s)/Host(s): R. Lafuente-Mercado

Mitigation of Applied Load Using Machine Learning for Adaptive Motor Control  

 

Kourosh Rahnamai, Jacob Rollins, Luke Moisan, Ryan Kayfus (Western New England Univ.)

An Analysis of the New EU AI Act and A Proposed Standardization Framework for Machine Learning Fairness  

Mike H.M. Teodorescu, Yongxu Sun, Haren N. Bhatia (Univ. of Washington), Christos Makridis (Univ. of Nicosia)

Reimagining Tacit Knowledge Extraction and Summarization by Inferencing Large Language Model in a High-Performance Computing Platform  

 

Rahul Hardikar, Saurabh Barve, Shampa Sarkar, Revati Kulkarni (TCS)

Data-Driven Dynamic Algorithm Dispatch with Large Language Models [Outstanding Short Paper Award]  

 

Rushil Shah, Emmanuel Lujan, Rabab Alomairy, Alan Edelman (MIT)

4-2: High Performance Computing Session (12:30-13:45)

Co-Chairs: D. Cousins & L. Zaidenberg

Easy Acceleration with Distributed Arrays

Jeremy Kepner, Chanup Byun, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel Burrill, Vijay Gadepally, Ryan Haney, Michael Houle, Matthew Hubbell, Hayden Jananthan,Michael Jones, Piotr Luszczek, Lauren Milechin, Guillermo Morales, Julie Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Peter Michaleas (MIT)
Performance Evaluation of LAPACK Using SVE Optimized BLAS Kernels

Aniket P. Garade, Sushil Singh, Vishal Rayala, Deepika H. V., Haribabu P., S. A. Kumar, S. D. Sudarsan (C-DAC)
Predicting HPC Job Run time with Realistic Data Using Application Input Parameters

Kenneth Lamar (Univ. of Central Florida), Benjamin A. Allan, M. Scot Swan, James M. Brandt (SNL), Damian Dechev (Univ. of Central Florida)
Generalized Methodology for Determining Numerical Features of Hardware Floating-Point Matrix Multipliers: Part I

Faizan A. Khattak, Mantas Mikaitis (Leeds Univ.)
Balancing Performance and Productivity: A Comparative Study of Apache Arrow vs. MPI

Ritvik Prabhu, Wu-chun Feng (Virginia Tech)

4-P2 (13:45-14:45): AI/ML/GenAI Poster Session

Chair(s)/Host(s): S.Mohindra

Centralized vs. Decentralized Security for Space AI Systems? A New Look  

Noam Schmitt, Marc Lacoste (Orange)
Improving Uncertainty Based Dataset Pruning with Density Estimation for Noisy Edge Environments

Mike Soricelli, Yuchou Chang (UMass Dartmouth), Christopher J Hixenbaugh (NUWC)

Virtual Benchmarking for HPC Systems Using ExaDigiT and Calculon  

 

Srishti Kalepu (Georgia Inst. of Tech.), Wesley Brewer, Matthias Maiterth (ORNL), Richard Vuduc (Georgia Inst. of Tech.)
An Improved Izhikevich Neuron Model with Time Delay and Adaptive Parameters for High-Performance Spiking Neural Networks

Alissa Kane (UMass Dartmouth), Felipe Marcelino, Anton Spirkin (NUWC), Yuchou Chang (UMass Dartmouth)
Adaptive Policy Synchronization for Scalable Reinforcement Learning

Rodney Lafuente Mercado (MIT Lincoln Laboratory)
LLMs in Crisis Triage: Benchmarking Zero-Shot Classification of Social Media

Emma L. McDaniel, Samuel Scheele, Jeffrey Liu (MIT Lincoln Laboratory)

4-3: High Performance Computing Session (14:15-15:30)

Co-Chairs: D. Cousins & P. Moniticciolo

Leveraging Caliper and Benchpark to Analyze MPI Communication Patterns: Insights from AMG2023, Kripke, and Laghos

Grace Nansamba, Evelyn Namugwanya (Tennessee Tech), David Boehme, Dewi Yokelson (LLNL), Riley Shipley (Tennessee Tech), Derek Schafer (Univ. of New Mexico), Michael McKinsey, Olga Pearce (LLNL), Anthony Skjellum (Tennessee Tech)
Employing High-Performance PETSc Network Simulation for Business Profit Analysis

Abu Asaduzzaman, Nowshin Nawal (Wichita State Univ.)
Performance Analysis of Inline Compression in pySDC [Outstanding Paper Award]

Emily Lattanzio (Clemson Univ.), Sansriti Ranjan (Siemens EDA), Robert Underwood (Argonne National Laboratory), Thomas Baumann, Robert Speck (Jülich Supercomputing Centre), Jon C. Calhoun (Clemson Univ.)
The NorthPole Validator: A Cycle-Accurate Simulator for HW/SW Codesign of a Prescheduled Neural Inference Accelerator

Alexander Andreopoulos, Michael V. Debole, Jeffrey A. Kusnitz, Nathaniel J. McClatchey, Tapan K. Nayak, Daniel F. Smith, Brian Taba, Filipp Akopyan, Rathinakumar Appuswamy, John V. Arthur, Andrew S. Cassidy, Pallab Datta, Carlos Ortega Otero, William P. Risk, Jun Sawada, Myron D. Flickner, Dharmendra S. Modha (IBM)

4-4: AI/ML/GenAI Session (15:45-17:15)

Co-Chairs: S. Mohindra & P. Moniticciolo

Evaluating Efficiency and Novelty of LLM-Generated Code for Graph Analysis [Outstanding Student Paper Award]

Atieh Barati Nia, Mohammad Dindoost, David A. Bader (New Jersey Inst. of Tech.)
Enhancing Sentiment Classification of E-commerce Reviews for Actionable Insights using LLMs and NLP

Kevin Power, Peter Harding, Jose Lopez, Alex Carroll, Elenna Dugundji (MIT)
DNN-Driven Task Scheduling for High Performance Edge-Cloud Heterogeneous Systems

Md Raihan Uddin, Abu Asaduzzaman, Fairuz Nawar, Christian Thompson (Wichita State Univ.)
Automating Harmonized System (HS) Code Classification from Unstructured Shipping Manifests using Large Language Models

Thomas Koch, Kevin Power (MIT)
A Time-Aware Sliding Window-Based Hotel Recommendation Framework Using Multi-Stage BERT-MRC

Md. Nazirul Hasan Shawon, K. M. Azharul Hasan (Khulna Univ. of Engr. and Tech.)
BanglaDocAtlas: A Multi-Class Annotated Dataset for Complex Bangla Document Layout Analysis

Md Safayat Hossain, Jannatul Ferdous (United Intl. Univ.), Md Raihan Uddin (Wichita State Univ.), K. M. A. Hossain, M. I. Ahmed, M. A. Rahman (United Intl. Univ.), Asif Sushmit (Bengali.AI), Farig Sadeque, Swakkhar Shatabda (BRAC University), Abu Asaduzzaman (Wichita State Univ.)

4-5: GenAI Opportunities & AI Challenges Session (17:30-19:30)

Organizers: V. Gadepally, D. Burrill & C. Prothmann

MI300A vs H100, LLM Concerns for Applications at NRL, and MCP Testing
William “Connor” Horne (NRL)

Trace Replay Simulation of MIT SuperCloud Dataset for Studying Optimal Policies for Sustainability [Outstanding Short Paper Award] 

 

Wesley Brewer, Matthias Maiterth (ORNL), Damien Fay (HPE)

Introduction: Advancing AI Challenges for the United States Department of the Air Force

Christian Prothmann, Vijay Gadepally, Jeremy Kepner, Koley Borchard, Luca Carlone, Zachary Folcik,J. Daniel Griffith, Michael Houle, Jonathan P. How, Nathan Hughes, Ifueko Igbinedion, Hayden Jananthan, Tejas Jayashankar, Michael Jones, Sertac Karaman, Binoy G. Kurien, Alejandro Lancho, Giovanni Lavezzi, Gary C. F. Lee, Charles E. Leiserson, Richard Linares, Lindsey McEvoy, Peter Michaleas, Chasen Milner, Alex Pentland, Yury Polyanskiy, Jovan Popovich, Jeffrey Price, Tim W. Reid, Stephanie Riley, Siddharth Samsi, Peter Saunders, Olga Simek, Mark S. Veillette, Amir Weiss, Gregory W. Wornell, Daniela Rus, Scott T. Ruppel (MIT)

Data-Driven Radio-Frequency Signal Separation Challenge

Yury Polyanskiy (MIT)
Tornado Network Challenge

Mark Veillette (MIT Lincoln Laboratory), Peter Saunders (USAF)
AI Innovation in Space Challenge

Giovanni Lavezzi, Richard Linares (MIT)
Evaluating Long-Context LLM Architectures with Controlled Synthetic Benchmarks Challenge

Olga Simek (MIT Lincoln Laboratory)
Predicting LLM Inference Server Request Capacity

Daniel J. Burrill, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alexander Bonn, Chansup Byun, Michael Houle, Matthew Hubbell, Michael Jones, Piotr Luszczek, Peter Michaleas, Guillermo Morales, Julie Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee (MIT Lincoln Laboratory), Vijay Gadepally (MIT)

Friday, September 19

5-K: Keynote Session (10:30-11:00)

Co-Chairs: J. Kepner & A. Reuther

Keynote Talk: Performance Engineering with MPI

Bill Gropp (NCSA)

5-1: Mixed Precision Session (11:00-12:15)

Co-Chairs: P. Luszczek & C. Byun

Invited Talk: Is Mixed Precision Computing really the Top Priority?

Hartwig Anzt (TU Munich)
Invited Talk: Reducing Numerical Precision Requirements in Quantum Chemistry Calculations

William Dawson (RIKEN)
Invited Talk: Why Exact Dot Products Obviate the Need for Mixed Precision

John Gustafson (ASU)
Invited Talk: Mixed Feelings about Mixed Precision

Hatem Ltaief (KAUST)
Invited Talk: Mixed Precision or Mixed Storage? User-Guided Compiler Transformations to Change the Data Layout On-the-Fly

Tobias Weinzierl (Durham Univ.)

5-2: Mixed Precision Session (12:30-13:45)

Co-Chairs: P. Luszczek & R. Muri

Invited Talk: Emulation Method for Matrix Multiplication

Katsuhisa Ozaki (Shibaura Inst. of Tech.)
A variable-precision implementation of the ADER-DG algorithm

Marc Marot-Lassauzaie, Michael Bader (Tech. Univ. Munich)
GPU-Accelerated, Mixed Precision GMRES(m) with Varied Restarts [Outstanding Student Paper Award]

Abir Haque, Suzanne Shontz, Xuemin Tu (Univ. of Kansas)
Performance and Numerical Aspects of Decompositional Factorizations with FP64 Floating-Point Emulation in INT8

Piotr Luszczek, Vijay Gadepally, LaToya Anderson, William Arcand, David Bestor, William Bergeron, Alex Bonn, Daniel J. Burrill, Chansup Byun, Michael Houle, Matthew Hubbell, Michael Jones, Peter Michaleas, Guillermo Morales, Julia Mullen, Andrew Prout, Albert Reuther, Antonio Rosa, Charles Yee, Jeremy Kepner (MIT Lincoln Laboratory)
Adaptive Spectral Block Floating Point for Discontinuous Galerkin Methods

Shivam Sundriyal, Markus Büttner (Univ. of Bayreuth), Christoph Alt, Tobias Kenter (Paderborn Univ.), Vadym Aizinger (Univ. of Bayreuth)

5-P2 (13:45-14:45): Embedded & GPU Poster Session

Chair(s)/Host(s): S. Shankar

Decision Support for Sustainable Agriculture Using I2C Sensors  

 

Anita Esmaeilian, Kishwar Ahmed (Univ. of Toledo)

ORB Feature Extraction on Embedded Platforms: A Heterogeneous CPU-GPU-PVA Approach  

 

Hamid Moghadaspour (Univ. of Coimbra), Nuno Neves (Univ. of Lisbon), Oscar Ferraz, Gabriel Falcao (Univ. of Coimbra)
Accelerating Temporal Triangle Counting and Betweenness Centrality on GPUs

Tuteja Trimansingh Parvindersingh, Venkata Kalyan Tavva (IIT Ropar), Subhasis Banerjee, Chiranjib Sur (Shell India Markets Pvt. Ltd.)
SCoDa: Scalable Community Detection in Data Streams

Akanksha Dwivedi, Prashant Srivastav, Dip Sankar Banerjee (IIT Jodhpur)
CTAM Tool for Hyperscaler Qualification

Tommy Yan, Rajat Madhusudan, Vani Pulendra, Anna Mary Mathew (Silicon Cloud HIE)

5-3: Cyber Analysis and Secure Computing Session (14:15-15:30)

Co-Chairs: D. Cousins & R. Vuduc

Invited Talk: Private Database Analytics with PAC Privacy

Srini Devadas (MIT)
Accelerating Multi-Party Computation Using Heterogeneous Systems [Outstanding Student Paper Award]

Xiteng Yao, Shining Yang, Mayank Varia, Martin Herbordt (Boston Univ.)
Optimizing Local Computation in Secure Matrix Multiplication for Outsourced Neural Networks [Outstanding Paper Award]

J. Parker Diamond, Andrea Lin, R. Nicholas Cunningham (MIT Lincoln Laboratory), Soamar Homsi (AFRL), John Darby Mitchell, Aseemit Pandey, Emily Shen (MIT Lincoln Laboratory)
A Framework For The Iterative Solution of Sparse Linear Systems on Hybrid Architectures Using Homomorphic Encryption

Lior Horesh, Vasileios Kalantzis, Barry M. Trager, Shashanka Ubaru (IBM Research)
Secure Virtual Network Embedding Through Fully Homomorphic Encryption

David Bruce Cousins, Carlo Pascoe (Duality Tech.), Erik Kline (USC)

5-4: Embedded Computing Session (15:45-17:15)

Co-Chairs: D. Cousins & E. Schnetzer

Invited Talk: AI and Biodiversity

Sara Beery (MIT)
On the Adaptation of Mixed-Radix Fast Fourier Transform for Resource-Constrained Environments [Outstanding Student Paper Award]

Atharva Gondhalekar, Paul Sathre, Wu-chun Feng (Virginia Tech)
Comparative Analysis of RISC-V Softcore and Hardcore Processors for Space Computing [Outstanding Student Paper Award]

Ni Nyoman Dhinar Gayatri, Alan D. George (Univ. of Pittsburgh)
TRACER Software Switching at Waveform Timescales

Connor Imes, Dong In D. Kang, Matthew French, John Paul Walters (USC Information Sciences Institute)
Hardware-Accelerated Transformer Framework for Real-Time Battery SoH Estimation

Talha Coskun (Univ. of Illinois Urbana-Champaign), Hiruna Vishwamith (Univ. of Moratuwa), Murat Isik (Stanford Univ.), I. Can Dikmen (Istinye Univ.)
A Qualifiable GPU Sharing Approach for AI Workloads in Critical Systems

Marc Solé i Bonet, Jannis Wolf (Barcelona Supercomputing Ctr.), Aridane Álvarez Suárez (fentISS), Leonidas Kosmidis (Barcelona Supercomputing Ctr.)

5-5: AI for Performance Engineering Session (17:30-19:30)

Organizers: H. Nguyen & D. Burrill

UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC [Outstanding Student Paper Award]

Tomer Bitan (Technion), Tal Kadosh (Ben-Gurion Univ. / IAEC), Erel Kaplan, Shira Meiri (Technion), Le Chen (Argonne NL), Peter Morales, Niranjan Hasabnis (Code Metal), Gal Oren (Stanford Univ.)
Web-Based Intelligent Decision Support System for Real-Time Toll Plaza Management and AI-Driven Operational Optimization

Pattarapon Klaykul, Wilaiporn Lee, Kanabadee Srisomboon, Luepol Pipanmekaporn, Akara Prayote (King Mongkut’s Univ.)
CRAMP: Categorizing Classifiers and Regressors for Scalable Parallelism on Distributed and Multicore Systems

Baidya Nath Saha, Pavan Sarvaiya, Wali Mohammad Abdullah, Md. Morshedul Islam (Concordia Univ. of Edmonton)
RAILS: Retrieval-Augmented Intelligence for Learning Software Development

Wali Mohammad Abdullah, Md. Morshedul Islam, Devraj Parmar, Happy Hasmukhbhai Patel, Sindhuja Prabhakaran, Baidya Saha (Concordia Univ. of Edmonton)
Towards Automated Reasoning Chains for Verification of LLM-Generated Scientific Code

Quentin Oschatz, Naifeng Zhang (Carnegie Mellon Univ.), Mike Franusich (SpiralGen), Franz Franchetti (Carnegie Mellon Univ.)
P4OMP: Retrieval-Augmented Prompting for OpenMP Parallelism in Serial Code

Wali Mohammad Abdullah (Concordia Univ. of Edmonton), Azmain Kabir (Univ. of Manitoba)
Towards -OmL: A Deep Learning Based Approach to Outperform Compiler Defaults

Hafsah Shahzad (Boston Univ.), Ahmed Sanaullah, Sanjay Arora, Ulrich Drepper (Red Hat), Martin Herbordt (Boston Univ.)

IEEE HPEC 2026