Close

Session

Research Manuscript
:
Architectures for Sparse, Adaptive, and Scalable Acceleration
DescriptionModern accelerator-centric systems must integrate specialized compute engines with efficient memory hierarchies and scalable interconnects. This session presents advances spanning sparse tensor compilation and streaming accelerators, runtime-adaptive FPGA and heterogeneous platforms, and innovations in memory and buffering, from tile-level optimizations and fine-grained cache partitioning to HBM performance characterization. The session concludes with work on distributed buffering and scale-up network infrastructure, highlighting how architectural techniques across compute, memory, and system levels collectively enable high-performance, adaptable, and scalable acceleration.
Event Type
Research Manuscript
TimeMonday, July 273:30pm - 5:30pm PDT
LocationMtg Room 202C
Topics
Design
Tracks
DES1-I. SoC, Heterogeneous, and Reconfigurable Architectures
Presentations
3:30pm - 3:42pm PDTSpArC: Sparse Tensor Accelerator Compilation with Scheduling and Mapping
3:42pm - 3:54pm PDTTRIDENT: An End-to-End Streaming Accelerator for TriSpGEMM
3:54pm - 4:06pm PDTProcyon: Promoting Fine-Grained Multi-Tenancy to Optimize Sparse Streaming Accelerators
4:06pm - 4:18pm PDTTRINE: A Token-Aware, Runtime-Adaptive FPGA Inference Engine for Multimodal AI
4:18pm - 4:30pm PDTPHAROS: Pipelined Heterogeneous Accelerators for Real-Time Safety-Critical Systems With Deadline Compliance
4:30pm - 4:42pm PDTCan Asymmetric Tile Buffering Be Beneficial?
4:42pm - 4:54pm PDTCachence: Fine-Grained Cache Partitioning in Both Time and Space
4:54pm - 5:06pm PDTToki: Profiling HBM Performance on FPGA Systems with RISC-V Soft Cores and PCIe Host DMA Traffic
5:06pm - 5:18pm PDTCredix: A Credit-Driven Distributed Buffer Management for Large-Scale Switching Chips
5:18pm - 5:30pm PDTOpenSUN: An Open Platform for Exploring Scale-Up Network Systems