Published on: September 2026
EMERGENT REASONING IN LARGE LANGUAGE MODELS: A SYSTEMATIC EVALUATION ACROSS TASK COMPLEXITY
N John Kuotsu
Article Status
Available Documents
Abstract
Keywords—large language models; emergent abilities; chain-of-thought reasoning; task complexity; benchmark evaluation; compositional generalization.
How to Cite this Paper
Kuotsu, N. J. (2026). Emergent Reasoning in Large Language Models: A Systematic Evaluation Across Task Complexity. International Journal of Creative and Open Research in Engineering and Management, <i>02</i>(9), 1-9. https://doi.org/10.55041/ijcope.v2i9.144
Kuotsu, N. "Emergent Reasoning in Large Language Models: A Systematic Evaluation Across Task Complexity." International Journal of Creative and Open Research in Engineering and Management, vol. 02, no. 9, 2026, pp. 1-9. doi:https://doi.org/10.55041/ijcope.v2i9.144.
Kuotsu, N. "Emergent Reasoning in Large Language Models: A Systematic Evaluation Across Task Complexity." International Journal of Creative and Open Research in Engineering and Management 02, no. 9 (2026): 1-9. https://doi.org/https://doi.org/10.55041/ijcope.v2i9.144.
References
[1] Wei et al., “Emergent abilities of large language models,” Trans. Mach. Learn. Res., 2022. doi: 10.48550/arXiv.2206.07682.[2] Kaplan et al., “Scaling laws for neural language models,” arXiv:2001.08361, 2020. doi: 10.48550/arXiv.2001.08361.
[3] Wei et al., “Chain-of-thought prompting elicits reasoning in large language models,” in Proc. Adv. Neural Inf. Process. Syst. (NeurIPS), 2022. doi: 10.48550/arXiv.2201.11903.
[4] Schaeffer, B. Miranda, and S. Koyejo, “Are emergent abilities of large language models a mirage?,” in Proc. Adv. Neural Inf. Process. Syst. (NeurIPS), 2023. doi: 10.48550/arXiv.2304.15004.
[5] B. Brown et al., “Language models are few-shot learners,” in Proc. Adv. Neural Inf. Process. Syst. (NeurIPS), 2020. doi: 10.48550/arXiv.2005.14165.
[6] Chowdhery et al., “PaLM: Scaling language modeling with pathways,” arXiv:2204.02311, 2022. doi: 10.48550/arXiv.2204.02311.
[7] Touvron et al., “LLaMA: Open and efficient foundation language models,” arXiv:2302.13971, 2023. doi: 10.48550/arXiv.2302.13971.
[8] Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa, “Large language models are zero-shot reasoners,” in Proc. Adv. Neural Inf. Process. Syst. (NeurIPS), 2022. doi: 10.48550/arXiv.2205.11916.
[9] Zhou et al., “Least-to-most prompting enables complex reasoning in large language models,” in Proc. Int. Conf. Learn. Represent. (ICLR), 2023. doi: 10.48550/arXiv.2205.10625.
Yao et al., “Tree of thoughts: Deliberate problem solving with large language models,” in Proc. Adv. Neural Inf. Process. Syst. (NeurIPS), 2023. doi: 10.48550/arXiv.2305.10601.
Ethical Compliance & Review Process
- •All submissions are screened under plagiarism detection.
- •Review follows editorial policy.
- •Authors retain copyright.
- •Peer Review Type: Double-Blind Peer Review
- •Published on: Sep 19 2026
This article is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License. You are free to share and adapt this work for non-commercial purposes with proper attribution.

