Fetching the paper…
Reading the bibliography…
Large Reasoning Models (LRMs) achieve superior performance by extending the thought length.
Training verifiers to solve math word problems, 2021
Cobbe, K., Kosaraju, V., Bavarian, M., Chen, M., Jun, H., Kaiser, L., Plappert, M., Tworek, J., Hilton, J., Nakano, R., Hesse, C., and Schulman, J · 2021
Earlier work this paper cites.
Let’s verify step by step, 2023
Lightman, H., Kosaraju, V., Burda, Y., Edwards, H., Baker, B., Lee, T., Leike, J., Schulman, J., Sutskever, I., and Cobbe, K · 2023
Earlier work this paper cites.
Compressed chain of thought: Efficient reasoning through dense representations, December 2024
Cheng, J. and Durme, B. V · 2024
Earlier work this paper cites.
Break the chain: Large language models can be shortcut reasoners, June 2024
Ding, M., Liu, H., Fu, Z., Song, J., Xie, W., and Zhang, Y · 2024
Earlier work this paper cites.
Training large language models to reason in a continuous latent space, December 2024
Hao, S., Sukhbaatar, S., Su, D., Li, X., Hu, Z., Weston, J., and Tian, Y · 2024
Earlier work this paper cites.
Olympiadbench: A challenging benchmark for promoting agi with olympiad-level bilingual multimodal scientific problems
He, C., Luo, R., Bai, Y., Hu, S., Thai, Z., Shen, J., Hu, J., Han, X., Huang, Y., Zhang, Y., Liu, J., Qi, L., Liu, Z., and Sun, M · 2024
Earlier work this paper cites.
Minicpm: Unveiling the potential of small language models with scalable training strategies
Hu, S., Tu, Y., Han, X., He, C., Cui, G., Long, X., Zheng, Z., Fang, Y., Huang, Y., Zhao, W., et al · 2024
Earlier work this paper cites.
Longllmlingua: Accelerating and enhancing llms in long context scenarios via prompt compression
Jiang, H., Wu, Q., Luo, X., Li, D., Lin, C.-Y., Yang, Y., and Qiu, L · 2024
Earlier work this paper cites.
OpenAI, :, Jaech, A., Kalai, A., Lerer, A., Richardson, A., El-Kishky, A., Low, A., Helyar, A., Madry, A., Beutel, A., Carney, A., Iftimie, A., Karpenko, A., Passos, A. T., Neitz, A., Prokofiev, A., Wei, A., Tam, A., Bennett, A., Kumar, A., Saraiva, A., Vallone, A., Duberstein, A., Kondrich, A., Mishchenko, A., Applebaum, A., Jiang, A., Nair, A., Zoph, B., Ghorbani, B., Rossen, B., Sokolowsky, B., Barak, B., McGrew, B., Minaiev, B., Hao, B., Baker, B., Houghton, B., McKinzie, B., Eastman, B., Lugaresi, C., Bassin, C., Hudson, C., Li, C. M., de Bourcy, C., Voss, C., Shen, C., Zhang, C., Koch, C., Orsinger, C., Hesse, C., Fischer, C., Chan, C., Roberts, D., Kappler, D., Levy, D., Selsam, D., Dohan, D., Farhi, D., Mely, D., Robinson, D., Tsipras, D., Li, D., Oprica, D., Freeman, E., Zhang, E., Wong, E., Proehl, E., Cheung, E., Mitchell, E., Wallace, E., Ritter, E., Mays, E., Wang, F., Such, F. P., Raso, F., Leoni, F., Tsimpourlas, F., Song, F., von Lohmann, F., Sulit, F., Salmon, G., Parascandolo, G., Chabot, G., Zhao, G., Brockman, G., Leclerc, G., Salman, H., Bao, H., Sheng, H., Andrin, H., Bagherinezhad, H., Ren, H., Lightman, H., Chung, H. W., Kivlichan, I., O’Connell, I., Osband, I., Gilaberte, I. C., Akkaya, I., Kostrikov, I., Sutskever, I., Kofman, I., Pachocki, J., Lennon, J., Wei, J., Harb, J., Twore, J., Feng, J., Yu, J., Weng, J., Tang, J., Yu, J., Candela, J. Q., Palermo, J., Parish, J., Heidecke, J., Hallman, J., Rizzo, J., Gordon, J., Uesato, J., Ward, J., Huizinga, J., Wang, J., Chen, K., Xiao, K., Singhal, K., Nguyen, K., Cobbe, K., Shi, K., Wood, K., Rimbach, K., Gu-Lemberg, K., Liu, K., Lu, K., Stone, K., Yu, K., Ahmad, L., Yang, L., Liu, L., Maksin, L., Ho, L., Fedus, L., Weng, L., Li, L., McCallum, L., Held, L., Kuhn, L., Kondraciuk, L., Kaiser, L., Metz, L., Boyd, M., Trebacz, M., Joglekar, M., Chen, M., Tintor, M., Meyer, M., Jones, M., Kaufer, M., Schwarzer, M., Shah, M., Yatbaz, M., Guan, M. Y., Xu, M., Yan, M., Glaese, M., Chen, M., Lampe, M., Malek, M., Wang, M., Fradin, M., McClay, M., Pavlov, M., Wang, M., Wang, M., Murati, M., Bavarian, M., Rohaninejad, M., McAleese, N., Chowdhury, N., Chowdhury, N., Ryder, N., Tezak, N., Brown, N., Nachum, O., Boiko, O., Murk, O., Watkins, O., Chao, P., Ashbourne, P., Izmailov, P., Zhokhov, P., Dias, R., Arora, R., Lin, R., Lopes, R. G., Gaon, R., Miyara, R., Leike, R., Hwang, R., Garg, R., Brown, R., James, R., Shu, R., Cheu, R., Greene, R., Jain, S., Altman, S., Toizer, S., Toyer, S., Miserendino, S., Agarwal, S., Hernandez, S., Baker, S., McKinney, S., Yan, S., Zhao, S., Hu, S., Santurkar, S., Chaudhuri, S. R., Zhang, S., Fu, S., Papay, S., Lin, S., Balaji, S., Sanjeev, S., Sidor, S., Broda, T., Clark, A., Wang, T., Gordon, T., Sanders, T., Patwardhan, T., Sottiaux, T., Degry, T., Dimson, T., Zheng, T., Garipov, T., Stasi, T., Bansal, T., Creech, T., Peterson, T., Eloundou, T., Qi, V., Kosaraju, V., Monaco, V., Pong, V., Fomenko, V., Zheng, W., Zhou, W., McCabe, W., Zaremba, W., Dubois, Y., Lu, Y., Chen, Y., Cha, Y., Bai, Y., He, Y., Zhang, Y., Wang, Y., Shao, Z., and Li, Z · 2024
The benefits of a concise chain of thought on problem-solving in large language models
Renze, M. and Guven, E · 2024
Cited alongside, same era.
L1: Controlling how long a reasoning model thinks with reinforcement learning, March 2025
Aggarwal, P. and Welleck, S · 2025
Cited alongside, same era.
Smollm2: When smol goes big – data-centric training of a small language model, 2025
Allal, L. B., Lozhkov, A., Bakouch, E., Blázquez, G. M., Penedo, G., Tunstall, L., Marafioti, A., Kydlíček, H., Lajarín, A. P., Srivastav, V., Lochner, J., Fahlgren, C., Nguyen, X.-S., Fourrier, C., Burtenshaw, B., Larcher, H., Zhao, H., Zakka, C., Morlon, M., Raffel, C., von Werra, L., and Wolf, T · 2025
Cited alongside, same era.
American mathematics competitions (amc), 2025
AMC · 2025
Cited alongside, same era.
Sketch-of-thought: Efficient llm reasoning with adaptive cognitive-inspired sketching, March 2025
Aytes, S. A., Baek, J., and Hwang, S. J · 2025
Cited alongside, same era.
S1: Simple test-time scaling, February 2025
Muennighoff, N., Yang, Z., Shi, W., Li, X. L., Fei-Fei, L., Hajishirzi, H., Zettlemoyer, L., Liang, P., Candès, E., and Hashimoto, T · 2025
Closest in time.
Optimizing test-time compute via meta reinforcement fine-tuning, March 2025
Qu, Y., Yang, M. Y. R., Setlur, A., Tunstall, L., Beeching, E. E., Salakhutdinov, R., and Kumar, A · 2025
Closest in time.
Qwq-32b: Embracing the power of reinforcement learning, March 2025
Qwen Team · 2025
Closest in time.
Codi: Compressing chain-of-thought into continuous space via self-distillation, February 2025
Shen, Z., Yan, H., Zhang, L., Hu, Z., Du, Y., and He, Y · 2025
Closest in time.
Token assorted: Mixing latent and text tokens for improved language model reasoning, February 2025
Su, D., Zhu, H., Xu, Y., Jiao, J., Tian, Y., and Zheng, Q · 2025
Closest in time.
Tokenskip: Controllable chain-of-thought compression in llms, February 2025
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Earlier work this paper cites.
Do not think that much for 2+3=? on the overthinking of o1-like llms, 2025
Chen, X., Xu, J., Liang, T., He, Z., Pang, J., Yu, D., Song, L., Liu, Q., Zhou, M., Zhang, Z., Wang, R., Tu, Z., Mi, H., and Yu, D · 2025
Cited alongside, same era.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, January 2025
DeepSeek-AI · 2025
Cited alongside, same era.
Token-budget-aware llm reasoning, February 2025
Han, T., Wang, Z., Fang, C., Zhao, S., Ma, S., and Chen, Z · 2025
Cited alongside, same era.
Retro-search: Exploring untaken paths for deeper and efficient reasoning, April 2025
Lu, X., Han, S., Acuna, D., Kim, H., Jung, J., Prabhumoye, S., Muennighoff, N., Patwary, M., Shoeybi, M., Catanzaro, B., and Choi, Y · 2025
Cited alongside, same era.
How well do llms compress their own chain-of-thought? a token complexity approach, March 2025a
Lee, A., Che, E., and Peng, T
Cited in the paper.
How well do llms compress their own chain-of-thought? a token complexity approach, 2025b
Lee, A., Che, E., and Peng, T
Cited in the paper.
Reasoning models can be effective without thinking
Ma, W., He, J., Snell, C., Griggs, T., Min, S., and Zaharia, M
Cited in the paper.
Xia, H., Li, Y., Leong, C. T., Wang, W., and Li, W · 2025
Closest in time.
Can atomic step decomposition enhance the self-structured reasoning of multimodal large models?, March 2025
Xiang, K., Liu, Z., Jiang, Z., Nie, Y., Cai, K., Yin, Y., Huang, R., Fan, H., Li, H., Huang, W., Zeng, Y., Yuan, Y.-J., Han, J., Hong, L., Xu, H., and Liang, X · 2025
Closest in time.
Chain of draft: Thinking faster by writing less, March 2025
Xu, S., Xie, W., Zhao, L., and He, P · 2025
Closest in time.
Lightthinker: Thinking step-by-step compression, February 2025
Zhang, J., Zhu, Y., Sun, M., Luo, Y., Qiao, S., Du, L., Zheng, D., Chen, H., and Zhang, N · 2025
Closest in time.