| 1 |
Bobby Yan |
The Right Kernel Every Time: Adaptive Sparse Compilation for Deep Learning |
| 2 |
Rubens Lacouture |
Programming Systems for Sparse Machine Learning on Modern Hardware |
| 3 |
Alexander Root |
Compiling Portable and Performant Spatial Queries with Data and Schedule Independence |
| 4 |
Sai Gautham Ravipati |
Portable Dynamic Tiling for Sparse Tensor Applications |
| 5 |
Atharva Chougule |
Partitioning Unstructured Sparse Tensor Algebra for Load-Balanced Parallel Execution |
| 6 |
Usman Tariq |
Fusion and Tiling for Computation Graphs |
| 7 |
Shiv Sundram |
Cyclotron: Specifying Dataflow in Distributed Architectures |
| 8 |
Chris Gyurgyik |
Data Layout Optimizations via Rewrite Rules |
| 9 |
Anderson Truong |
Flexible analytical performance modeling for deep learning on hierarchical, heterogeneous hardware |
| 10 |
Yasmine Omri |
Agent Memory: System Implications and Co-design for Multi-Tenant Long-Horizon Workloads |
| 11 |
Michael Oduoza |
HW/SW Co-design for SSMs at the Edge |
| 12 |
Allen Pan |
Extending Voyager for Digital Compute-in-Memory-Based Accelerator Generation and Compilation |
| 13 |
Bo Wun Cheng |
Interplay of Activation Quantization and Sparsification for LLM Compression |
| 14 |
Wonsuk Jang |
SemanticDialect: Semantic-Aware Mixed-Format Quantization for Video Diffusion Transformers |
| 15 |
Jeffrey Yu |
Sphinx: An Edge-LLM Accelerator with Outlier-Aware W4A4 Microscaling and Fused 2b KV-Cache Decoding |
| 16 |
Christian Kubicka |
uVLA: Multi-Chiplet 16nm SoP with HMoP for Physical AI |
| 17 |
Yuchen Mei |
Agate: A Design Space Exploration System for Heterogeneous CGRAs Accelerating Dense ML Workloads |
| 18 |
Po-Han Chen |
SPEC: A Scalable and Physical-Design-Friendly Architecture for CGRAs |
| 19 |
Áron Ricardo Perez-Lopez |
Pono 2.0: A Versatile SMT-Based Model Checker for Safety and Liveness |
| 20 |
Zhouhua Xie |
Provably Correct Low-Precision Approximations of Nonlinear Functions for AI Hardware |
| 21 |
Elizaveta Pertseva |
Automated Geometric Predicate Synthesis |
| 22 |
Ritvik Sharma |
CoTenN: Constrained Optimization with Tensor Networks |