Please note that ISTA Research Explorer no longer supports Internet Explorer versions 8 or 9 (or earlier).

We recommend upgrading to the latest Internet Explorer, Google Chrome, or Firefox.

1 Publication


2025 | Published | Conference Paper | IST-REx-ID: 19877 | OA
Frantar E, Castro RL, Chen J, Hoefler T, Alistarh D-A. 2025. MARLIN: Mixed-precision auto-regressive parallel inference on Large Language Models. Proceedings of the 30th ACM SIGPLAN Annual Symposium on Principles and Practice of Parallel Programming. PPoPP: Symposium on Principles and Practice of Parallel Programming, 239–251.
[Published Version] View | Files available | DOI | arXiv
 

Filters and Search Terms

isbn=9798400714436

Search

Filter Publications

  • Display / Sort

    Citation Style: ISTA Annual Report

    Export / Embed