BEGIN:VCALENDAR
VERSION:2.0
PRODID:Linklings LLC
BEGIN:VTIMEZONE
TZID:America/Chicago
X-LIC-LOCATION:America/Chicago
BEGIN:DAYLIGHT
TZOFFSETFROM:-0600
TZOFFSETTO:-0500
TZNAME:CDT
DTSTART:19700308T020000
RRULE:FREQ=YEARLY;BYMONTH=3;BYDAY=2SU
END:DAYLIGHT
BEGIN:STANDARD
TZOFFSETFROM:-0500
TZOFFSETTO:-0600
TZNAME:CST
DTSTART:19701101T020000
RRULE:FREQ=YEARLY;BYMONTH=11;BYDAY=1SU
END:STANDARD
END:VTIMEZONE
BEGIN:VEVENT
DTSTAMP:20260202T201803Z
LOCATION:260
DTSTART;TZID=America/Chicago:20251116T113000
DTEND;TZID=America/Chicago:20251116T120000
UID:submissions.supercomputing.org_SC25_sess206_ws_exampi102@linklings.com
SUMMARY:Large-Message All-to-All Communication at Frontier Scale
DESCRIPTION:James B. White III (Oak Ridge National Laboratory (ORNL), Nati
 onal Center for Computational Sciences)\n\nNear the full scale of exascale
  supercomputers, latency can dominate the cost of all-to-all communication
  even for very large message sizes. We describe GPU-aware all-to-all imple
 mentations designed to reduce latency for large message sizes at extreme s
 cales, and we present their performance using 65536 tasks (8192 nodes) on 
 the Frontier supercomputer at the Oak Ridge Leadership Computing Facility.
  Two implementations perform best for different ranges of message size, an
 d all outperform the vendor-provided MPI_Alltoall. Our results show promis
 ing options for improving implementations of MPI_Alltoall_init.\n\nRecordi
 ng: Livestreamed, Recorded\n\nRegistration Category: Technical Program Reg
  Pass, Workshop Reg Pass\n\nSession Chairs: Matthew G. F. Dosanjh (Sandia 
 National Laboratories); William Schonbein (Sandia National Laboratories); 
 Amanda J. Bienz (University of New Mexico); and Joseph Schuchart (Stony Br
 ook University, Institute for Advanced Computational Science (IACS))\n\n
END:VEVENT
END:VCALENDAR
