Talking to machines
Prompt engineering as an epistemic competence
DOI:
https://doi.org/10.54103/2531-5994/31801Keywords:
Ingegneria dei prompt, Modello di ragionamento, Gioco di ruolo, LLM, Apprendimento contestuale, Alfabetizzazione critica nell'IA, IA agenticaAbstract
This paper proposes to regard prompt engineering not as a repertoire of techniques for improving the performance of generative language models, but as an epistemic competence that defines the communicative interaction between humans and artificial intelligence systems. Having identified its conditions of possibility in in-context learning and in the fictional personae generated by LLMs, it turns to a brief account of the classical prompting techniques, from few-shot to chain-of-thought. The advent of reasoning models and agentic architectures absorbs many of these formalized strategies into the models themselves; what is transformed, without disappearing, is the competence required to construct epistemically controlled interactions. Prompt engineering is thereby repositioned within a broader critical competence, in which asking pertinent questions, verifying sources and assessing the reliability of the output matters more than proficiency in codified protocols.
Downloads
References
Alammar, J., & Grootendorst, M. (2024). Hands-on large language models: Language understanding and generation (1st ed.). O'Reilly Media.
Anthropic. (2025). Effective context engineering for AI agents. https://www.anthropic.com/engineering/effective-context-engineering-for-ai-agents
Barez, F., Wu, T.-Y., Arcuschin, I., Lan, M., Wang, V., Siegel, N., Collignon, N., Neo, C., Lee, I., Paren, A., Bibi, A., Trager, R., Fornasiere, D., Yan, J., Elazar, Y., & Bengio, Y. (2025). Chain-of-thought is not explainability [Preprint]. alphaXiv. https://www.alphaxiv.org/abs/2025.02v1
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D. M., Wu, J., Winter, C., … Amodei, D. (2020). Language models are few-shot learners. Advances in Neural Information Processing Systems, 33, 1877–1901.
Chalmers, D. J. (2026). What we talk to when we talk to language models [Preprint]. PhilArchive. https://philarchive.org/rec/CHAWWT-8
Chen, R., Arditi, A., Sleight, H., Evans, O., & Lindsey, J. (2025). Persona vectors: Monitoring and controlling character traits in language models [Preprint]. arXiv. https://doi.org/10.48550/arXiv.2507.21509
Chen, Y., Benton, J., Radhakrishnan, A., Uesato, J., Denison, C., Schulman, J., Somani, A., Hase, P., Wagner, M., Roger, F., Mikulik, V., Bowman, S. R., Leike, J., Kaplan, J., & Perez, E. (2025). Reasoning models don't always say what they think [Preprint]. arXiv. https://doi.org/10.48550/arXiv.2505.05410
Clark, A., & Chalmers, D. (1998). The extended mind. Analysis, 58(1), 7–19. https://doi.org/10.1093/analys/58.1.7 DOI: https://doi.org/10.1093/analys/58.1.7
Dai, D., Sun, Y., Dong, L., Hao, Y., Ma, S., Sui, Z., & Wei, F. (2023). Why can GPT learn in-context? Language models secretly perform gradient descent as meta-optimizers. In A. Rogers, J. Boyd-Graber, & N. Okazaki (Eds.), Findings of the Association for Computational Linguistics: ACL 2023 (pp. 4005–4019). Association for Computational Linguistics. https://doi.org/10.18653/v1/2023.findings-acl.247 DOI: https://doi.org/10.18653/v1/2023.findings-acl.247
Dell'Acqua, F., McFowland, E., III, Mollick, E. R., Lifshitz-Assaf, H., Kellogg, K., Rajendran, S., Krayer, L., Candelon, F., & Lakhani, K. R. (2023). Navigating the jagged technological frontier: Field experimental evidence of the effects of AI on knowledge worker productivity and quality (Working Paper No. 24-013). Harvard Business School. https://doi.org/10.2139/ssrn.4573321 DOI: https://doi.org/10.2139/ssrn.4573321
Devlin, J., Chang, M.-W., Lee, K., & Toutanova, K. (2019). BERT: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) (pp. 4171–4186). Association for Computational Linguistics. https://doi.org/10.18653/v1/N19-1423 DOI: https://doi.org/10.18653/v1/N19-1423
Feng, G., Zhang, B., Gu, Y., Ye, H., He, D., & Wang, L. (2023). Towards revealing the mystery behind chain of thought: A theoretical perspective. In Advances in Neural Information Processing Systems (Vol. 36, pp. 70757–70798). Curran Associates. DOI: https://doi.org/10.52202/075280-3100
Greshake, K., Abdelnabi, S., Mishra, S., Endres, C., Holz, T., & Fritz, M. (2023). Not what you've signed up for: Compromising real-world LLM-integrated applications with indirect prompt injection. In Proceedings of the 16th ACM Workshop on Artificial Intelligence and Security (AISec '23) (pp. 79–90). Association for Computing Machinery. https://doi.org/10.1145/3605764.3623985 DOI: https://doi.org/10.1145/3605764.3623985
Huang, X., Liu, W., Chen, X., Wang, X., Wang, H., Lian, D., Wang, Y., Tang, R., & Chen, E. (2024). Understanding the planning of LLM agents: A survey [Preprint]. arXiv. https://doi.org/10.48550/arXiv.2402.02716
Janus. (2022, September 2). Simulators. LessWrong. https://www.lesswrong.com/posts/vJFdjigzmcXMhNTsx/simulators
Kambhampati, S. (2024). Can large language models reason and plan? Annals of the New York Academy of Sciences, 1534(1), 15–18. https://doi.org/10.1111/nyas.15125 DOI: https://doi.org/10.1111/nyas.15125
Karpathy, A. [@karpathy]. (2025, June 25). +1 for "context engineering" over "prompt engineering". People associate prompts with short task descriptions you'd give an LLM in your day-to-day use. When in every industrial-strength LLM app, context engineering is the delicate art and science of filling the context window [Post]. X. https://x.com/karpathy/status/1937902205765607626
Kojima, T., Gu, S. S., Reid, M., Matsuo, Y., & Iwasawa, Y. (2022). Large language models are zero-shot reasoners. In Advances in Neural Information Processing Systems (Vol. 35, pp. 22199–22213). Curran Associates. DOI: https://doi.org/10.52202/068431-1613
Liu, J., Shen, D., Zhang, Y., Dolan, B., Carin, L., & Chen, W. (2022). What makes good in-context examples for GPT-3? In Proceedings of Deep Learning Inside Out (DeeLIO 2022): The 3rd Workshop on Knowledge Extraction and Integration for Deep Learning Architectures (pp. 100–114). Association for Computational Linguistics. https://doi.org/10.18653/v1/2022.deelio-1.10 DOI: https://doi.org/10.18653/v1/2022.deelio-1.10
Lu, Y., Bartolo, M., Moore, A., Riedel, S., & Stenetorp, P. (2022). Fantastically ordered prompts and where to find them: Overcoming few-shot prompt order sensitivity. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) (pp. 8086–8098). Association for Computational Linguistics. https://doi.org/10.18653/v1/2022.acl-long.556 DOI: https://doi.org/10.18653/v1/2022.acl-long.556
Marks, S., Lindsey, J., & Olah, C. (2026, February 23). The persona selection model: Why AI assistants might behave like humans. Anthropic Alignment Science Blog. https://www.anthropic.com/research/persona-selection-model
Mei, L., Yao, J., Ge, Y., Wang, Y., Bi, B., Cai, Y., Liu, J., Li, M., Li, Z.-Z., Zhang, D., Zhou, C., Mao, J., Xia, T., Guo, J., & Liu, S. (2025). A survey of context engineering for large language models [Preprint]. arXiv. https://doi.org/10.48550/arXiv.2507.13334
Meincke, L., Mollick, E. R., Mollick, L., & Shapiro, D. (2025). Prompting science report 1: Prompt engineering is complicated and contingent [Working paper]. SSRN. https://doi.org/10.2139/ssrn.5165270 DOI: https://doi.org/10.2139/ssrn.5165270
Min, S., Lyu, X., Holtzman, A., Artetxe, M., Lewis, M., Hajishirzi, H., & Zettlemoyer, L. (2022). Rethinking the role of demonstrations: What makes in-context learning work? In Y. Goldberg, Z. Kozareva, & Y. Zhang (Eds.), Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing (pp. 11048–11064). Association for Computational Linguistics. https://doi.org/10.18653/v1/2022.emnlp-main.759 DOI: https://doi.org/10.18653/v1/2022.emnlp-main.759
Mollick, E. (2024). Co-intelligence: Living and working with AI. Portfolio/Penguin.
Noy, S., & Zhang, W. (2023). Experimental evidence on the productivity effects of generative artificial intelligence. Science, 381(6654), 187–192. https://doi.org/10.1126/science.adh2586 DOI: https://doi.org/10.1126/science.adh2586
OpenAI. (2025). Reasoning best practices. OpenAI Platform Documentation. Retrieved June 1, 2026, from https://platform.openai.com/docs/guides/reasoning-best-practices
Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A., Schulman, J., Hilton, J., Kelton, F., Miller, L., Simens, M., Askell, A., Welinder, P., Christiano, P., Leike, J., & Lowe, R. (2022). Training language models to follow instructions with human feedback. In Advances in Neural Information Processing Systems (Vol. 35, pp. 27730–27744). Curran Associates. DOI: https://doi.org/10.52202/068431-2011
Pan, J., Gao, T., Chen, H., & Chen, D. (2023). What in-context learning "learns" in-context: Disentangling task recognition and task learning. In A. Rogers, J. Boyd-Graber, & N. Okazaki (Eds.), Findings of the Association for Computational Linguistics: ACL 2023 (pp. 8298–8319). Association for Computational Linguistics. https://doi.org/10.18653/v1/2023.findings-acl.527 DOI: https://doi.org/10.18653/v1/2023.findings-acl.527
Roncaglia, G. (2023). L'architetto e l'oracolo: Forme digitali del sapere da Wikipedia a ChatGPT. Laterza.
Shanahan, M. (2024). Still "talking about large language models": Some clarifications [Preprint]. arXiv. https://doi.org/10.48550/arXiv.2412.10291
Shanahan, M., McDonell, K., & Reynolds, L. (2023). Role play with large language models. Nature, 623(7987), 493–498. https://doi.org/10.1038/s41586-023-06647-8 DOI: https://doi.org/10.1038/s41586-023-06647-8
Sprague, Z., Yin, F., Rodriguez, J. D., Jiang, D., Wadhwa, M., Singhal, P., Zhao, X., Ye, X., Mahowald, K., & Durrett, G. (2025). To CoT or not to CoT? Chain-of-thought helps mainly on math and symbolic reasoning [Conference paper]. In The Thirteenth International Conference on Learning Representations (ICLR 2025). https://doi.org/10.48550/arXiv.2409.12183
Templeton, A., Conerly, T., Marcus, J., Lindsey, J., Bricken, T., Chen, B., Pearce, A., Citro, C., Ameisen, E., Jones, A., Cunningham, H., Turner, N. L., McDougall, C., MacDiarmid, M., Freeman, C. D., Sumers, T. R., Rees, E., Batson, J., Jermyn, A., … Henighan, T. (2024). Scaling monosemanticity: Extracting interpretable features from Claude 3 Sonnet. Transformer Circuits Thread. https://transformer-circuits.pub/2024/scaling-monosemanticity/
Turpin, M., Michael, J., Perez, E., & Bowman, S. R. (2023). Language models don't always say what they think: Unfaithful explanations in chain-of-thought prompting. In A. Oh, T. Naumann, A. Globerson, K. Saenko, M. Hardt, & S. Levine (Eds.), Advances in Neural Information Processing Systems (Vol. 36, pp. 74952–74965). Curran Associates. DOI: https://doi.org/10.52202/075280-3275
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin, I. (2017). Attention is all you need. In Advances in Neural Information Processing Systems (Vol. 30, pp. 5998–6008). Curran Associates.
Von Oswald, J., Niklasson, E., Randazzo, E., Sacramento, J., Mordvintsev, A., Zhmoginov, A., & Vladymyrov, M. (2023). Transformers learn in-context by gradient descent. In A. Krause, E. Brunskill, K. Cho, B. Engelhardt, S. Sabato, & J. Scarlett (Eds.), Proceedings of the 40th International Conference on Machine Learning (pp. 35151–35174). PMLR. https://proceedings.mlr.press/v202/von-oswald23a.html
Wang, L., Ma, C., Feng, X., Zhang, Z., Yang, H., Zhang, J., Chen, Z., Tang, J., Chen, X., Lin, Y., Zhao, W. X., Wei, Z., & Wen, J. (2024). A survey on large language model based autonomous agents. Frontiers of Computer Science, 18(6), Article 186345. https://doi.org/10.1007/s11704-024-40231-1 DOI: https://doi.org/10.1007/s11704-024-40231-1
Wang, X., Wei, J., Schuurmans, D., Le, Q., Chi, E., Narang, S., Chowdhery, A., & Zhou, D. (2023). Self-consistency improves chain of thought reasoning in language models [Conference paper]. In The Eleventh International Conference on Learning Representations (ICLR 2023). https://doi.org/10.48550/arXiv.2203.11171
Wang, Z., Mao, S., Wu, W., Ge, T., Wei, F., & Ji, H. (2024). Unleashing the emergent cognitive synergy in large language models: A task-solving agent through multi-persona self-collaboration. In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers) (pp. 257–279). Association for Computational Linguistics. https://doi.org/10.18653/v1/2024.naacl-long.15 DOI: https://doi.org/10.18653/v1/2024.naacl-long.15
Wei, J., Wang, X., Schuurmans, D., Bosma, M., Ichter, B., Xia, F., Chi, E., Le, Q., & Zhou, D. (2022). Chain-of-thought prompting elicits reasoning in large language models. In Advances in Neural Information Processing Systems (Vol. 35, pp. 24824–24837). Curran Associates. DOI: https://doi.org/10.52202/068431-1800
Xie, S. M., Raghunathan, A., Liang, P., & Ma, T. (2022). An explanation of in-context learning as implicit Bayesian inference [Conference paper]. In The Tenth International Conference on Learning Representations (ICLR 2022). https://doi.org/10.48550/arXiv.2111.02080
Yao, S., Yu, D., Zhao, J., Shafran, I., Griffiths, T. L., Cao, Y., & Narasimhan, K. (2023). Tree of thoughts: Deliberate problem solving with large language models. In Advances in Neural Information Processing Systems (Vol. 36, pp. 11809–11822). Curran Associates. DOI: https://doi.org/10.52202/075280-0517
Yao, S., Zhao, J., Yu, D., Du, N., Shafran, I., Narasimhan, K., & Cao, Y. (2023). ReAct: Synergizing reasoning and acting in language models [Conference paper]. In The Eleventh International Conference on Learning Representations (ICLR 2023). https://doi.org/10.48550/arXiv.2210.03629
Zheng, C., Liu, Z., Xie, E., Li, Z., & Li, Y. (2023). Progressive-hint prompting improves reasoning in large language models [Preprint]. arXiv. https://doi.org/10.48550/arXiv.2304.09797
Zheng, M., Pei, J., Logeswaran, L., Lee, M., & Jurgens, D. (2024). When "a helpful assistant" is not really helpful: Personas in system prompts do not improve performances of large language models. In Y. Al-Onaizan, M. Bansal, & Y.-N. Chen (Eds.), Findings of the Association for Computational Linguistics: EMNLP 2024 (pp. 15126–15154). Association for Computational Linguistics. https://doi.org/10.18653/v1/2024.findings-emnlp.888 DOI: https://doi.org/10.18653/v1/2024.findings-emnlp.888
Downloads
Published
How to Cite
Issue
Section
License
Copyright (c) 2026 Fabio Ciotti (Author)

This work is licensed under a Creative Commons Attribution-ShareAlike 4.0 International License.
Accepted 2026-07-16
Published 2026-08-06

