← Search

Pragaash Ponnusamy

3 accepted papers

2025

Training-Free Activation Sparsity in Large Language Models

ICLR 2025spotlight

Activation sparsity can enable practical inference speedups in large language models (LLMs) by reducing the compute and memory-movement required for matrix multiplications during the forward pass. However, existing methods face limitations that inhibit widespread adoption. Some approaches are tail…

2024

Mechanistic Design and Scaling of Hybrid Architectures

ICML 2024poster

The development of deep learning architectures is a resource-demanding process, due to a vast design space, long prototyping times, and high compute costs associated with at-scale model training and evaluation. We set out to simplify this process by grounding it in an end-to-end mechanistic architec…

2022

Self-Aware Feedback-Based Self-Learning in Large-Scale Conversational AI

NAACL 2022industry

Self-learning paradigms in large-scale conversational AI agents tend to leverage user feedback in bridging between what they say and what they mean. However, such learning, particularly in Markov-based query rewriting systems have far from addressed the impact of these models on future training wher…

Cited by 3SourcePDFScholar