The 2nd Workshop on Sparsity, the Key Ingredient from HPC to Efficient LLMs

Co-Located with MICRO 2026, Athens, Greece

Sunday, November 1, 2026 | 1:00 PM to 5:00 PM EET | Location: InterContinental Athenaeum Athens Hotel

Sparsity has become a defining feature in modern computing workloads, from scientific simulations on HPC platforms to inference and training in cutting-edge LLMs. It appears across all layers of the stack: bit-level computations, sparse data structures, irregular memory access patterns, high-level architectural design such as MoEs and dynamic routing, and system-level concerns including rack-scale deployment and scale-up/scale-out networking for massive models. Although sparsity offers enormous potential to improve computing efficiency, reduce energy consumption, and enable scalability, its integration into modern systems introduces significant architectural, algorithmic, and programming challenges. The SPICE workshop brings together the architecture, systems, HPC, and machine learning communities to explore the growing role of sparsity as a foundational tool for scaling efficiency and shaping the next generation of computing systems. By fostering collaboration among researchers, practitioners, and industry experts, the workshop will focus on (a) developing novel architectures and system techniques that exploit sparsity at multiple levels and (b) deploying sparsity-aware models effectively in real-world scientific and AI applications. The format includes keynotes from academic and industry leaders, peer-reviewed paper presentations, and interactive sessions for collaboration.

Keynote Talk

Photo of Josep Torrellas

Title TBA

by Josep Torrellas

Bio

Josep Torrellas is the Thomas M. Siebel Chair in Computer Science at the University of Illinois at Urbana-Champaign (UIUC). He is the Director of the ACE Center for Evolvable Computing (an SRC/DARPA JUMP 2.0 Center), past Co-Leader of an Intel Strategic Research Alliance (ISRA) on Computer Security, and past Director of the Illinois-Intel Parallelism Center (I2PC). His research interests are multiprocessor computer architectures and parallel computing. Some of his contributions include thread-level speculation (TLS) architectures, the Bulk Multiprocessor concept, deterministic record and replay mechanisms, process variation mitigation techniques, and hardware defenses against speculative execution attacks. In addition, he has contributed to several experimental multiprocessor designs such as IBM’s PERCS Multiprocessor, Intel’s Runnemede Extreme-Scale Multiprocessor, Illinois Cedar, and Stanford DASH. Among other awards, he has received the IEEE CS Harry H. Goode Memorial Award and the IEEE CS Edward J. McCluskey Technical Achievement Award. He is an IEEE Computer Society Golden Core Member, and a Fellow of IEEE, ACM, and AAAS. He received a PhD from Stanford University.

Abstract

Abstract TBA.

Call For Papers

We invite submissions that address any aspect of sparsity in computing systems. Topics of interest include, but are not limited to:

We welcome complete papers, early stage work, and position papers that inspire discussion and foster community building. We target a soft limit of 4 pages, formatted in double-column style, similar to the main MICRO submission. If you have any questions please feel free to reach out to Bahar Asgari [bahar at umd dot edu] or Ramyad Hadidi [rhadidi at d-matrix dot ai]

Important Info:

FAQ:

1. How strict is the 4-page limit?

The 4-page limit is a soft guideline. Your text (excluding references) may slightly exceed 4 pages (e.g., 4.25–4.5 pages). The exact length will not affect the decision on your paper.

2. Do early-stage or position papers need to include results?

Yes. Even early-stage or position papers should include some preliminary results to support their claims. We understand these papers may not yet have a complete set of evaluation results.

3. Should I list the authors in my submission?

No. In line with the main MICRO submission guidelines, please do not include author names in your submission.

4. Can I also submit my work elsewhere?

Yes. Papers submitted to SPICE will not be published in the proceedings. You are free to publish the complete version of your work elsewhere, and you may also submit preliminary or ongoing work to SPICE.

Organizers

Co-Chairs

Bahar Asgari Assistant Professor University of Maryland, College Park (UMD)
Ramyad Hadidi Senior Staff Engineer d-Matrix

Organizer Committee (Alphabetical Order)

Antonino Tumeo Chief Scientist Pacific Northwest National Laboratory (PNNL)
Ben Feinberg Senior Member of Technical Staff Sandia National Laboratories
Farzaneh Zokaee System-on-Chip Architect Samsung Research America (SRA)
Olivia Hsu Assistant Professor Carnegie Mellon University (CMU)
Prashant J. Nair Associate Professor / Senior Principal Engineer UBC & d-Matrix

Student Volunteers

Helia Hosseini Publicity Chair PhD Student University of Maryland, College Park (UMD)
Johnson Umeike Submission Chair PhD Student University of Maryland, College Park (UMD)