Media Summary: ... feature maps throughout the backbone to avoid deteriorating these features through repeated application of the Presented by Nitish Srivastava at HPCA2020, San Diego, CA, United States. Abstract: Tensor factorizations are powerful tools in ... This the full presentation of our work at

Hpca Spatten Efficient Sparse Attention - Detailed Analysis & Overview

... feature maps throughout the backbone to avoid deteriorating these features through repeated application of the Presented by Nitish Srivastava at HPCA2020, San Diego, CA, United States. Abstract: Tensor factorizations are powerful tools in ... This the full presentation of our work at Project & Seminar, ETH Zürich, Spring 2022 Hands-on Acceleration on Heterogeneous Computing Systems ... Dead Page and Dead Block Predictors: Cleaning TLBs and Caches Together Authors: Chandrashis Mazumdar*, Prachatos Mitra* ... This is a short presentation (lightning talk) of the paper entitled "Prodigy: Improving the Memory Latency of Data-Indirect Irregular ...

Photo Gallery

Short Intro HPCA'21 SpAtten: Efficient Sparse Attention Architecture with Cascade Token/Head Pruning
HPCA' SpAtten: Efficient Sparse Attention Architecture w/ Cascade Token/Head Pruning by Hanrui Wang
Chip Demo for "SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning"
Is Sparse Attention more Interpretable?
Arxiv 2021: Sparse attention Planning
[HPCA'20] Tensaurus: A Versatile Accelerator for Mixed Sparse-Dense Tensor Computations
F-1 Roofline Model (HPCA 2021)
Hanrui Wang's Talk at HPCA'20 on "SpArch: Efficient Architecture for Sparse Matrix Multiplication"
[HPCA' 21 Full] Stream Floating: Enabling Proactive and Decentralized Cache Optimizations
HPCA-2023 CTA: Hardware-Software Co-design for Compressed Token Attention Mechanism
HetSys Course: Lecture 10: Parallel Patterns: Sparse Matrices (Spring 2022)
[HPCA 2021, Main talk] Dead Page and Dead Block Predictors: Cleaning TLBs and Caches Together
View Detailed Profile
Short Intro HPCA'21 SpAtten: Efficient Sparse Attention Architecture with Cascade Token/Head Pruning

Short Intro HPCA'21 SpAtten: Efficient Sparse Attention Architecture with Cascade Token/Head Pruning

Short intro video for

HPCA' SpAtten: Efficient Sparse Attention Architecture w/ Cascade Token/Head Pruning by Hanrui Wang

HPCA' SpAtten: Efficient Sparse Attention Architecture w/ Cascade Token/Head Pruning by Hanrui Wang

Talk video for

Chip Demo for "SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning"

Chip Demo for "SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning"

We design and tape out the chip for

Is Sparse Attention more Interpretable?

Is Sparse Attention more Interpretable?

Video for ACL 2021 paper https://arxiv.org/abs/2106.01087.

Arxiv 2021: Sparse attention Planning

Arxiv 2021: Sparse attention Planning

... feature maps throughout the backbone to avoid deteriorating these features through repeated application of the

[HPCA'20] Tensaurus: A Versatile Accelerator for Mixed Sparse-Dense Tensor Computations

[HPCA'20] Tensaurus: A Versatile Accelerator for Mixed Sparse-Dense Tensor Computations

Presented by Nitish Srivastava at HPCA2020, San Diego, CA, United States. Abstract: Tensor factorizations are powerful tools in ...

F-1 Roofline Model (HPCA 2021)

F-1 Roofline Model (HPCA 2021)

Talk on F-1 Roofline model at

Hanrui Wang's Talk at HPCA'20 on "SpArch: Efficient Architecture for Sparse Matrix Multiplication"

Hanrui Wang's Talk at HPCA'20 on "SpArch: Efficient Architecture for Sparse Matrix Multiplication"

Presentation at

[HPCA' 21 Full] Stream Floating: Enabling Proactive and Decentralized Cache Optimizations

[HPCA' 21 Full] Stream Floating: Enabling Proactive and Decentralized Cache Optimizations

This the full presentation of our work at

HPCA-2023 CTA: Hardware-Software Co-design for Compressed Token Attention Mechanism

HPCA-2023 CTA: Hardware-Software Co-design for Compressed Token Attention Mechanism

The presentation video in

HetSys Course: Lecture 10: Parallel Patterns: Sparse Matrices (Spring 2022)

HetSys Course: Lecture 10: Parallel Patterns: Sparse Matrices (Spring 2022)

Project & Seminar, ETH Zürich, Spring 2022 Hands-on Acceleration on Heterogeneous Computing Systems ...

[HPCA 2021, Main talk] Dead Page and Dead Block Predictors: Cleaning TLBs and Caches Together

[HPCA 2021, Main talk] Dead Page and Dead Block Predictors: Cleaning TLBs and Caches Together

Dead Page and Dead Block Predictors: Cleaning TLBs and Caches Together Authors: Chandrashis Mazumdar*, Prachatos Mitra* ...

Prodigy - HPCA 2021 best paper award winner (lightning talk)

Prodigy - HPCA 2021 best paper award winner (lightning talk)

This is a short presentation (lightning talk) of the paper entitled "Prodigy: Improving the Memory Latency of Data-Indirect Irregular ...