Session Full Schedule · Contributors · Organizations · Search Program · My Schedule · Happening NowMore…Search ProgramMy ScheduleHappening NowResearch and ACM SRC Posters: Poster Presentations (Research, ACM SRC Grads/Undergrads)Session ChairsKento SatoRIKEN Center for Computational Science (R-CCS)Anja GerbesGeorg-August-Universität GöttingenChris SchlipaliusPawsey Supercomputing Research CentreCommonwealth Scientific and Industrial Research Organisation (CSIRO), AustraliaEvent TypeResearch and ACM SRC PostersTimeFriday, 21 November 20258:00am - 12:00pm CSTLocationSecond Floor AtriumTagsResearch & ACM SRC PostersRegistration Categories TP Similar SessionsBest Poster Finalist Presentations (ACM SRC Grads)Best Poster Finalist Presentations (ACM SRC Undergrads)Doctoral Showcase I PresentationsPresentationsIntelligent Surrogates Pay Attention to Data, Improving Multi-Objective HPC OptimizationAuthorsAshna Nawar AhmedBanooqa BandayTerry JonesTanzima Z. IslamNumerical Investigation of Radiation Hydrodynamic Instabilities at Scale with FleCSI-HARDAuthorsMåns I. AnderssonIsaac C. BannermanMoon B. HazarikaAkshit JariwalaJonathan MathurinMadela B. QuashieJulien LoiseauHyun LimHigh-Performance Sparse Attention on Tensor Cores: Fused3S and BeyondAuthorZitong LiA Quantum Solver for Multidimensional Partial Differential Equations: Practical Case StudiesAuthorsManu ChaudharyKareem El-ArabyAlvir NobelIshraq IslamManish SinghSunday OgundeleKieran EganSneha ThomasVincent VordtriedeDevon BontragerSerom KimEsam El-ArabyEvaluating the Power-Monitoring Capabilities of AuroraAuthorPrecious EyabiDiOMP-Offloading: Portable OpenMP Offloading for Distributed Heterogeneous SystemsAuthorsBaodi ShanMauricio Araya-PoloBarbara ChapmanPerformance Engineering of Scientific Applications with MVAPICH and TAU Using Emerging Communication PrimitivesAuthorsDhabaleswar K. (DK) PandaSameer ShendeAhmad AbdelfattahYifeng CuiEnhancing Usability and Performance in Experimental Environments ManagementAuthorsZahra TemoriPaul MarshallKate KeaheyCan Lossy Compression Benefit NVMe-Based I/O?AuthorsDarren NgDuo ZhangSheng DiZhaorui ZhangXiaoyi LuWONDERS: Integrating WOW, PONDER, and SCALE for Enhanced Scheduling PerformanceAuthorsFabian LehmannJonathan RauJonathan BaderOdej KaoUlf LeserTensor Core Accelerated Fast Multipole Method for GROMACSAuthorsJiamian HuangMuhammad Umair SadiqRio YokotaBerk HessEchoes of Earth: Building an Autonomous Environmental Lab for Acoustic SensingAuthorsHudson ReynoldsAlex TueckeMike ShermanKate KeaheyMitigating I/O Bottlenecks in LiDAR Pipelines by Directly Merging Neural Decompression and Semantic SegmentationAuthorsEthan MarquezMax FaykusOyinlolu OdetoyeMelissa SmithJon CalhounAccelerating Scientific Workflows with LLM-Driven Compiler Optimizations for Generated High-Performance HardwareAuthorsRobert RamstadNicolas Bohm AgostiniAntonino TumeoAutoSlim: Intelligent Automata Graph Optimization for Efficient AccelerationAuthorsTiffany YuRasha KarakchiJob Grouping-Based Intelligent Resource Recommendation FrameworkAuthorsBeste OztopBenjamin SchwallerVitus J. LeungJim BrandtBrian KulisManuel EgeleAyse K. CoskunParallel Local Motif Counting on Large-Scale Dynamic GraphsAuthorsAli KhanSanjukta BhowmickMichela TauferDiffPro: Joint Timestep and Layer-Wise Precision Optimization for Efficient Diffusion InferenceAuthorsFarhana AminKanchon GharamiDimitrios S. NikolopoulosScODA: An Emerging Pipeline for Evaluating Distributed Database Performance To Support Operational Data AnalyticsAuthorsNicholas SynovicFNU ShilpikaSilvio RizziDoug WaldronGeorge K. ThiruvathukalMichael E. PapkaMulti-GPU Implementation and Roofline Analysis of a Numerical Global Ocean ModelAuthorsTakateru YamagishiMasao KurogiTakao KawasakiYoshimasa MatsumuraHiroyasu HasumiHarmony: Converged Supercomputer Scratch and Archival FilesystemsAuthorJake CarrollPractical Viability of Translating Legacy Fortran Code to C++ Using Large Language ModelsAuthorsRen ImaiMasatoshi KawaiKeichi TakahashiHiroyuki TakizawaScaling Singular Values Beyond GPU Memory Limits: Out-of-Core, GPU-Accelerated, and Unified Across Data Precision and HardwareAuthorEvelyne RingootJulia with Intelligent Runtime for Heterogeneous ComputingAuthorsNarasinga Rao MiniskarPedro Valero-LaraWilliam GodoyKeita TeranishiJeffrey S. VetterForward Error Bounds and Efficient Algorithms for Computing a Tensor Times Matrix Chain in Low Precision on GPUsAuthorsJulian BellavitaPiyush SaoRamakrishnan KannanNovel Graph Alignment Algorithms for Identifying Non-Determinism in Large-Scale SimulationsAuthorDhroov PandeyHydraCache: LLM Inference Prefill Parallelization Through Distributed Cache BlendingAuthorsAdib Rezaei ShahmirzadiShayan ShabihiMona MoghadampanahFurong HuangDimitrios S. NikolopoulosEnergy-Efficient Multimodal LLM Inference: Stage-Level Characterization and Input-Aware ControlsAuthorsMona MoghadampanahAdib Rezaei ShahmirzadiDimitrios S. NikolopoulosLocal vs. Global FFT Approaches for High-Performance Ultrasound Simulation on Multi-GPU SystemsAuthorsOliver KuníkJiri JarosCUR-MoE: Portable Mixture-of-Experts with Interpretable High-Ratio CompressionAuthorRitesh BhirudTowards Application Agnostic HPC ProfilingAuthorsHari Teja JajulaDhruva KulkarniBrian AustinPurushotham BangaloreCompute System Simulator: Modeling the Impact of Allocation Policy and Hardware Reliability on HPC Cloud Resource UtilizationAuthorsJarrod LeddyHuseyin YildizMixed Compute Environments with OpenCHAMIAuthorsSean GibsonRichard KimSamuel QuanTravis CottonThomas MackellPhySiViT: A Physics Simulation Vision TransformerAuthorsJessica EzembaJames AffulMei-Yu WangTemplate Task-Based Multiresolution Analysis in Hybrid EnvironmentsAuthorsNilesh ChaturvediJoseph SchuchartRobert J. HarrisonScalable Alternative Route Computation with ACE: A C++17 Library for HPC Traffic SimulationsAuthorsPaulo SilvaPavlína SmolkováKateřina SlaninováJan MartinovičJoão BarbosaMatej ŠpeťkoEmanuele VitaliMojo: Python-Like MLIR-Based GPU Portable Science KernelsAuthorTatiana MelnichenkoReal-Time ML-Based Defense Against Malicious Payload in Reconfigurable Embedded SystemsAuthorsRye Stahle-SmithRasha KarakchiBridging the Quantum Coding Gap: Instruction-Tuned LLMs for QiskitAuthorsSixu ChenYuqi ZhangQiang GuanShipping HPC Ecosystems Across Platforms: Portable and Composable HPC Clusters as CodeAuthorsGerman Felipe Giraldo VillaThéo GrivelGeorge IoannidisEdita KizinevicCarolina LindqvistNicolas LitchinkoPablo LlopisAntonio Javier RussoGilles FouresteyConfiguring Large Language Models for Regional Ocean Model DevelopmentAuthorsAidan JanneyGiovanni Seijo-EllisDan AmrheinUsing Hardware Metrics To Understand Performance of the RAJA Performance Suite Kernels in Different GPU Modes on MI300AAuthorsAmr AbouelmagdStephanie BrinkMichael McKinseyDavid BoehmeJason BurmarkBrian RyujinTom ScoglandOlga PearceMemory-Efficient CFD Based on MPS: Effective One-Billion-Cell Resolution on a Single NodeAuthorsJunya OnishiAyato TakiiSangwon KimYounghwa ChoMakoto TsubokuraEnabling Real-Time, Extreme-Scale Bayesian Inference: FFT-Based GPU-Accelerated Matrix-Vector Products for Block-Triangular Toeplitz MatricesAuthorsSreeram VenkatOmar GhattasA Formal Characterization of Non-Monotonicity in Tensor CoresAuthorsPaul JiangVivian ZhengSync-Free GPU Parallelization of Sparse Kernels from Sequential Python CodeAuthorMalko-Bani SomoHeterogeneity-Aware Task Allocation for Modern HPC SystemsAuthorsSowmya YellapragadaJessica Imlau DagostiniKevin GottRebecca Hartman-BakerBuilding the Foundation for Machine Learning-Based Mars Weather ForecastingAuthorMohammad AltiwainyAccelerating AI Co-Scientists with HPC InfrastructureAuthorSuryatejas AppanaDivergence Prediction System for CFD SimulationsAuthorsTakashi SogaTakanori UchidaSusumu DateOptimizing the GPU All-Reduce Using Multiple Processes Per GPUAuthorsMichael AdamsAmanda BienzProcess-Based Predictors of Vulnerability ReintroductionAuthorsSamiha ShimmiNicholas SynovicMona RahimiGeorge ThiruvathukalShortcut Mixup Policy: Toward Improving Robustness and Speed in Goal-Conditioned RLAuthorsMatthew HyattYassir AtlasHal BryntesonDiego Roa PerdomoAthena AngaraMengjiao HanJoseph InsleyJanet KnowlesYongho KimVictor MateevitsiMichael PapkaSilvio RizziGeorge ThiruvathukalNicola FerrierOrchid: Towards Heterogeneous Batched Eigenvalue SolversAuthorMatthew ChungApplying Lossy Compression Techniques to GNN TrainingAuthorsMilan ShahReece NeffMichela BecchiCharacterizing Performance and Energy Trade-Offs on the Aurora SupercomputerAuthorsSolomon BekeleSwann PerarnauBrice VideauLearning To Select Scheduling Algorithms in OpenMPAuthorsJonas H. Müller KorndörferAli MohammedAhmed EleliemyQuentin GuilloteauReto KrummenacherFlorina CiorbaMassively Parallel Bayesian Inference Framework for GPU Supercomputers: Application to Estimation of Coseismic Fault SlipAuthorsKai NakaoTsuyoshi IchimuraKohei FujitaOptimizing Task-Driven Offloading in LLVMAuthorsJan KrausJoachim JenkeChristian TerbovenLuthier: A Dynamic Binary Instrumentation Framework Targeting AMD GPUsAuthorsMatin Raayai-ArdakaniNorman RubinDavid KaeliTidalMark: A Scalable Benchmark for Coastal Water Level ForecastingAuthorsLucas RaicuDaniel GrzendaIan FosterKyle ChardSeamless Scaling of Applications Across Programming ModelsAuthorsReto KrummenacherQuentin GuilloteauJonas H. Müller KorndörferFlorina M. CiorbaWafer-Scale Simulation of Mutator Allele Dynamics in Large Asexual PopulationsAuthorsMatthew Andres MorenoEmily DolsonLuis ZamanDistributed Modular Digital Twin Network for High-Performance and Reliable Data CentersAuthorsYan ChenXing LuCary FaulknerAlex VlachokostasHanlong WanJeremy LerondExploring Fine-Grained Parallelism in Data-Flow Runtime Systems on Many-Core SystemsAuthorsWenyi WangMaxime GonthierHaibin LaiPoornima NookalaHaochen PanIan FosterIoan RaicuKyle ChardAdversaGuard: A Distributed Data Poisoning Benchmark for Parallel AIAuthorsYulia KumarSolomon ThomasDejaun GayleJ. Jenny LiDov KrugerThe Impact of Maximum Vector Length on Cache Management Techniques in RISC-V Vector ExtensionAuthorsShunya NomuraJiaheng LiuKeichi TakahashiHiroyuki TakizawaCROSS-HPC System Bayesian Optimization with Adaptive TransferAuthorsAbrar HossainKishwar AhmedFacilitating Mixed Python-Fortran HPC Codes: 4D Drift-Kinetic Simulations with PyccelAuthorsEmily BourneYaman GüçlüWhen Label Propagation Outperforms BFS in Breadth-First Graph TraversalAuthorsKalsuda LapborisuthSrinivas AluruAnalyzing Dataset Popularity for Optimizing In-Network StorageAuthorsGunwoo KimAlex SimKesheng WuMassively Parallel GPU Rasterizer for Next-Generation Computational LithographyAuthorsLoay HegazyMohamed TaherSherif HammoudaExplicit Low-Order Finite-Element Wave Simulation Accelerated with Variable-Precision Computing Using INT8 Tensor CoresAuthorsKohei FujitaTsuyoshi IchimuraMuneo HoriLalith MaddegedaraWiCAT: Reducing Congestion at Wireless Interfaces in Heterogeneous ArchitecturesAuthorTarun SharmaMPI-SGX: Enabling Confidential Computing for MPI Parallel Applications with Intel SGX TechnologyAuthorsKota ShimojimaHayato YamakiHiroki HondaShinichiro MatsuoAtsuko TakefusaShinobu MiwaVaultX Merge: Breaking Memory Barriers in Proof-of-Space Plot GenerationAuthorsArnav SirigereVarvara BondarenkoIoan RaicuAn Approach for Correlating Compiler Optimizations with Runtime PerformanceAuthorsBefikir BogaleOlga PearceTom ScoglandMichela TauferIncineRate: Multi-Modal FPGA Accelerator for SCNNsAuthorsBjörn A. LindqvistArtur PodobasOptimizing and Extending Periodogram Computations for AstronomyAuthorsYuwei SunLehman GarrisonGNNs on Evolving Graphs: A Benchmark of Incremental Updates and Meta-Learning ApproachesAuthorsSriram SrinivasanSanjukta BhowmickHamdan AlabsiRand ObeidatDetecting Silent Data Corruption in Sparse Matrices Using Hardware Performance CountersAuthorsMinseop ChoiOrlando AriasSeung Woo SonInference-as-a-Service Prototype at NERSCAuthorsColin ThomasPo-Han HuangHilary UtaegbulamJohannes BlaschkeBruno CoimbraPengfei DingXiangyang JuAndrew NaylorMichael WangFrom Legacy to Portable: An Agentic AI Workflow for Fortran Code Translation and Cross-Architecture OptimizationAuthorsSparsh GuptaKamalavasan KamalakkannanMaxim MoraruGalen ShipmanPatrick DiehlUnmasking Performance Variability in GPU Codes on SupercomputersAuthorsCunyang WeiKeshav PradeepAbhinav BhateleLeveraging Large Language Models for Property Prediction in Polymorphic Organic SemiconductorsAuthorsShreya PagariaMei-Yu WangDana O’ConnorJulian UranPaola BuitragoScalable Execution Framework for R on Manycore SystemsAuthorsXiran ZhangJavier ConejeroSameh AbdulahJorge EjarqueYing SunRosa M. BadiaDavid E. KeyesMarc G. GentonC++ Standard Parallelism for GPU Programming in a Particle-In-Cell ApplicationAuthorsEster El KhouryMathieu LobetJulien BigotLaurent ColombetBetween the NIC and a Hard Place: Evaluating 400 Gb/s Ethernet for HPC Data TransfersAuthorsAdelle FerrisEvelyn NeedhamNikole GrandezJesse MartinezDoug EganHigh Performance Batch SVD Using GPUsAuthorAhmad AbdelfattahDistributed 3D Gaussian Splatting for High-Resolution Isosurface VisualizationAuthorsMengjiao HanAndres SewellJoseph InsleyJanet KnowlesVictor A. MateevitsiMichael E. PapkaSteve PetruzzaSilvio RizziGPU Kernels for Mixture of ExpertsAuthorsArthur FeeneyYing Wai LiAparna ChandramowlishwaranDivide, Conquer, and Denoise: Hybrid Parallel Diffusion with Memory-Aware Coarse-to-Fine InferenceAuthorsFarhana AminKanchon GharamiDimitrios NikolopoulosCATIOS: Time-Resolved I/O-Aware Job Scheduling for HPC SystemsAuthorsYuTsen TsengMasatoshi KawaiKeichi TakahashiHiroyuki TakizawaUnraveling Distant Galaxies: Analyzing IFU Data with Parsl and AcademyAuthorDaniel BabniggUnderstanding GPU Utilization Using LDMS Data on PerlmutterAuthorsOnur CankurBrian AustinAbhinav BhateleA Scalability Study of Quantum Algorithms for Dimensionality Reduction of Multidimensional DataAuthorsKareem El-ArabyThom PopovicAlvir NobelSunday OgundeleKatherine KlymkoDaan CampsAnastasiia ButkoEsam El-ArabyUnified Performance Modeling Stack for Distributed GPU Applications: Complementing Analytical Insights with Machine LearningAuthorUrvij SaroliyaFast Linear Solvers via AI-Tuned Markov Chain Monte Carlo-Based Matrix InversionAuthorsAnton LebedevWon Kyung LeeSoumyadip GhoshOlha I. YamanVassilis KalantzisYingdong LuTomasz NowickiShashanka UbaruLior HoreshVassil AlexandrovHardware-Aware Quantum Circuit SynthesisAuthorsNathan JonesAkhilesh BondapalliToby CoxIan LewisRong GeParaViz3D: MPI Trace Visualization with 3D VideoAuthorsJean-Yves VerhaegheGeorg HagerAyesha AfzalChameleon Concierge: Retrieval-Augmented Generation (RAG) To Enhance Open Testbed DocumentationAuthorSaieda Ali ZadaAdvancing EEG Signal Analysis with Quantum Machine LearningAuthorsStephanie MurrayErika ParsonsTowards a GPU-Accelerated Web-Based Graph Rendering Framework for Large-Scale Protein NetworksAuthorsJiaxin LuLandon DykenShilpika ShilpikaVenkatram VishwanathMichael PapkaSidharth KumarFrom Petabytes to Predictions: Harnessing Large-Scale NeuroBlu Mental Health Data and ML To Mitigate Medication Non-AdherenceAuthorsAlyson CollinsCathy SandovalMaya SeshanSrishti SrivastavaJosh McWilliamsOptimizing Collectives with Large Payloads on GPU-Based SupercomputersAuthorsSiddharth SinghMahua SinghKeshav PradeepAbhinav BhateleA Kokkos-Based Proxy of the Exascale Metagenome Assembler MetaHipMer2: A First Use of Kokkos for Computational BiologyAuthorsLogan WilliamsGavin ConantMichela BecchiJan CieskoAmy PowellUnderstanding Communication Bottlenecks in Multi-Node LLM InferenceAuthorsPrajwal SinghaniaSiddharth SinghLannie Dalton HoughIshan RevankarHarshitha MenonCharles JekelAbhinav BhateleEuropean Open Web Index: Large Complex Graph VisualizationAuthorsPavlina SmolkovaKaterina SlaninovaEvaluating the Usage of Python Libraries on a Production SupercomputerAuthorThomas PapkaCIRE: LLVM Analysis for Floating-Point Rounding Error Affected by Precision and OptimizationsAuthorsCayden LundTanmay TirpankarGanesh GopalakrishnancsDF: A Double-Float Arithmetic Library for the Cerebras CS-2AuthorsReo NagashimaAkeru NakamuraKai MurakamiRyunosuke MatsuzakiDaichi MukunokiTakaaki MiyajimaSRAP: Sender-Side Receiver-Aware Port Selection for High-Speed Multi-Flow TCPAuthorsShingo HattoriOsamu TatebeCan Long-Haul RDMA Benefit Federated Learning?AuthorsZhonghao ChenYuke LiDuo ZhangXiaoyi LuAlgorithms and Applications of Dynamic Network Analysis Using CANDYAuthorsAashish PandeyArindam KhandaS.M. ShovanAli Y. KhanBoyana NorrisSajal K. DasSanjukta BhowmickUnderstanding LLM Behavior on HPC Data via Mechanistic InterpretabilityAuthorsMd Mahbubur RahmanArjun GuhaHarshitha MenonChatHPC: Building the Foundations for a Productive and Trustworthy AI-Assisted HPC EcosystemAuthorsPedro Valero-LaraAaron YoungMohammad Alaul Haque MonilSwaroop PophaleZheming JinJeffrey S. VetterKeita TeranishiWilliam F. GodoyProductive Scalable Distributed Task Scheduling Using an MPI-based Backend for DaggerAuthorYan GuimarãesEvaluating LiDAR Compression for 3D Semantic Segmentation in Diverse Off-Road Environments on GOOSE DatasetAuthorsAdam NiemczuraMax FaykusOyinlolu OdetoyeMelissa SmithJon CalhounScott GroelTime-Stepping Hamiltonian Simulation for Solving Nonlinear PDEs via a Quantum-Classical Hybrid ApproachAuthorsSangwon KimJunya OnishiAyato TakiiYounghwa ChoTsubokura MakotoAn Agent-Based Viral Venture: Adaptive Tool Selection for Scalable GenomicsAuthorsNaomi KolodisnerAlok Kamatar (Advisor)J. Greg Pauloski (Advisor)Accelerating Linear Solve with Mixed Precision Nested Recursive Subdivision on AI HardwareAuthorVicki CarricaScalable Multi-Node Multi-GPU Datalog Engine with Energy-Aware ProfilingAuthorsAhmedur Rahman ShovonSidharth KumarEnabling Efficient Runtime Data Analysis to a Crystal Deformation SimulationAuthorArthur JaquardA Toolbox for Load Balancing Development and Analysis in WarpX/AMReX ApplicationsAuthorsJessica Imlau DagostiniSowmya YellapragadaKevin GottRebecca Hartman-BakerAn Efficient GEMM Acceleration Method for LLM Inference with Variable-Length SequencesAuthorsYu ZhangLu LuClassifying Performance Bounds Using Machine LearningAuthorsLewis LittmanTom DeakinGATSched: Multi-Objective Graph Attention Networks for Energy-Efficient HPC Job SchedulingAuthorKyrian AdimoraRange Search on Heterogeneous Systems with Processing-in-Memory ArchitectureAuthorsTasmia JannatSatish PuriMichael GowanlockJACC: Easy CPU/GPU Performance Portability for Scientific Applications in JuliaAuthorsWilliam GodoyPedro Valero-LaraPhilip FacklerKeita TeranishiJeffrey VetterJhonny GonzalezJose GonzalezAlexis Huante