FR-Spec
-
Last Updated 19 June, 2026
-
by David Spuler, Ph.D.
FR-Spec: Book Excerpts and Blog Articles
Free online book excerpts with full text chapters online and free PDF downloads, and the Aussie AI blog, including related articles:
- David Spuler, May 31st, 2026, Chapter 49. Eagle, Medusa, and FR-Spec, in book LLM Inference Optimization: State-of-the-Art Research, Table of Contents: https://www.aussieai.com/book/llm-inference-optimization https://www.amazon.com/dp/B0H3FKR39T
Research on FR-Spec
Research papers include:
- Miles Williams, Young D. Kwon, Rui Li, Alexandros Kouris, Stylianos I. Venieris, 14 Feb 2026, Speculative Decoding with a Speculative Vocabulary, https://arxiv.org/abs/2602.13836 (Adaptively per-token reducing the drafter's vocabulary for a smaller unembedding matrix.)
- Weilin Zhao, Tengyu Pan, Xu Han, Yudi Zhang, Ao Sun, Yuxiang Huang, Kaihuo Zhang, Weilun Zhao, Yuxuan Li, Jianyong Wang, Zhiyuan Liu, Maosong Sun, 20 Feb 2025, FR-Spec: Accelerating Large-Vocabulary Language Models via Frequency-Ranked Speculative Sampling, https://arxiv.org/abs/2502.14856, https://github.com/thunlp/FR-Spec (Limiting the draft model in speculative decoding to frequently-used tokens.)
- Raghavv Goel, Sudhanshu Agrawal, Mukul Gagrani, Junyoung Park, Yifan Zao, He Zhang, Tian Liu, Yiping Yang, Xin Yuan, Jiuyan Lu, Chris Lott, Mingu Lee, 3 Jul 2025 (v2), VOCABTRIM: Vocabulary Pruning for Efficient Speculative Decoding in LLMs, https://arxiv.org/abs/2506.22694
- Wilhelm Tranheden, Shahnawaz Ahmed, Devdatt Dubhashi, Jonna Matthiesen, Hannes von Essen, 15 Mar 2026, FlashHead: Efficient Drop-In Replacement for the Classification Head in Language Model Inference, https://arxiv.org/abs/2603.14591
- Jinbin Zhang, Nasib Ullah, Erik Schultheis, Rohit Babbar, 3 Feb 2026 (v3), DynaSpec: Context-aware Dynamic Speculative Sampling for Large-Vocabulary Language Models, https://arxiv.org/abs/2510.13847
- Nadav Timor, Jonathan Mamou, Oren Pereg, Hongyang Zhang, David Harel, 2 Jun 2025, Out-of-Vocabulary Sampling Boosts Speculative Decoding, https://arxiv.org/abs/2506.03206
- David Spuler, May 31st, 2026, Chapter 49. Eagle, Medusa, and FR-Spec, in book LLM Inference Optimization: State-of-the-Art Research, Table of Contents: https://www.aussieai.com/book/llm-inference-optimization https://www.amazon.com/dp/B0H3FKR39T
- Shuyu Zhang, Lingfeng Pan, Qicheng Wang, Yaqi Shi, Yueyang Tan, Ruyu Yan, Jiaqi Chen, Lixing Du, Lu Wang, 28 May 2026 (v2), EvoSpec: Evolving Speculative Decoding via Real-Time Vocabulary and Parameter Adaptation, https://arxiv.org/abs/2605.27390
More AI Research Topics
Read more about:
- 500+ LLM Inference Optimization Techniques
- What's Hot in LLM Inference Optimization in 2025?
- Inference Optimization Research
- « Research Home