> For the complete documentation index, see [llms.txt](https://paper.lingyunyang.com/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://paper.lingyunyang.com/reading-notes/conference/icml-2023.md).

# ICML 2023

## Meta Info

Homepage: <https://icml.cc/Conferences/2022>

Paper List: <https://icml.cc/virtual/2023/papers.html?filter=titles>

## Papers

### LLM Inference

* Deja Vu: Contextual Sparsity for Efficient LLMs at Inference Time \[[Paper](https://proceedings.mlr.press/v202/liu23am.html)]
* FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU \[[Personal Notes](/reading-notes/miscellaneous/arxiv/2023/flexgen.md)] \[[Paper](https://proceedings.mlr.press/v202/sheng23a.html)]
