Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
LG AI Research blog post summarizing the LayerSkip method from Meta presented at ACL 2024, which uses layer dropout, early exit loss, and self-speculative decoding to speed up LLM inference.
From the source
In this blog, we will explore various approaches to enhancing the efficiency of large language models (LLMs), with a focus on research presented at ACL 2024.
lgresearch.ai