Feigenbaum, E. A. The art of artificial intelligence: themes and case studies of knowledge engineering. In Proc. 5th International Joint Conference on Artificial Intelligence (IJCAI ’77) Vol. 2, 1014–1029 (Morgan Kaufmann, 1977).Thank you for reading this post, don’t forget to subscribe! Russell, S. J. & Norvig, P. Artificial Intelligence: A Modern Approach (Prentice Hall, 1995).
Feigenbaum, E. A. The art of artificial intelligence: themes and case studies of knowledge engineering. In Proc. 5th International Joint Conference on Artificial Intelligence (IJCAI ’77) Vol. 2, 1014–1029 (Morgan Kaufmann, 1977).
Thank you for reading this post, don't forget to subscribe!Russell, S. J. & Norvig, P. Artificial Intelligence: A Modern Approach (Prentice Hall, 1995).
Sutton, R. S. & Barto, A. Reinforcement Learning: An Introduction (MIT Press, 1998).
Wooldridge, M. in Multiagent Systems: A Modern Approach to Distributed Artificial Intelligence (ed. Weiss, G.) 27–79 (MIT Press, 1999).
Sumers, T. R., Yao, S., Narasimhan, K. & Griffiths, T. L. Cognitive architectures for language agents. In Trans. Machine Learning Research (2024).
Yee, L., Chui, M. & Roberts, R. Why agents are the next frontier of generative AI. McKinsey & Company https://www.mckinsey.com/capabilities/mckinsey-digital/our-insights/why-agents-are-the-next-frontier-of-generative-ai/ (2024).
Salesforce. Salesforce unveils Agentforce–what AI was meant to be. Salesforce https://www.salesforce.com/uk/news/press-releases/2024/09/12/agentforce-announcement/ (2024).
Spataro, J. New autonomous agents scale your team like never before. The Official Microsoft Blog https://blogs.microsoft.com/blog/2024/10/21/new-autonomous-agents-scale-your-team-like-never-before/ (2024).
Casper, S. et al. The AI agent index. Preprint at https://doi.org/10.48550/arXiv.2502.01635 (2025).
Gabriel, I. et al. The ethics of advanced AI assistants. Preprint at https://doi.org/10.48550/arXiv.2404.16244 (2024).
Kirk, H. R., Gabriel, I., Summerfield, C., Vidgen, B. & Hale, S. A. Why human–AI relationships need socioaffective alignment. Humanit. Soc. Sci. Commun. 12, 728 (2025).
Google Scholar
Chan, A. et al. Harms from increasingly agentic algorithmic systems. In Proc. 2023 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’23) 651–666 (Association for Computing Machinery, 2023).
Uuk, R. et al. A taxonomy of systemic risks from general-purpose AI. Preprint at https://doi.org/10.48550/arXiv.2412.07780 (2024).
Kasirzadeh, A. Two types of AI existential risk: decisive and accumulative. Philos. Stud. 182, 1975–2003 (2025).
Google Scholar
Weidinger, L. et al. Taxonomy of risks posed by language models. In Proc. 2022 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’22) 214–229 (Association for Computing Machinery, 2022).
Bird, C., Ungless, E., & Kasirzadeh, A. Typology of risks of generative text-to-image models. In Proc. 2023 AAAI/ACM Conference on AI, Ethics, and Society (AIES ’23) 396–410 (Association for Computing Machinery, 2023).
Hammond, L. et al. Multi-agent risks from advanced AI. Preprint at https://doi.org/10.48550/arXiv.2502.14143 (2025).
Kolt, N. Governing AI agents. Notre Dame Law Rev. https://doi.org/10.2139/ssrn.4772956 (2025).
Kraprayoon, J., Williams, Z., & Fayyaz, R. AI agent governance: a field guide. Preprint at https://doi.org/10.48550/arXiv.2505.21808 (2025).
Chan, A. et al. Visibility into AI agents. In Proc. 2024 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’24) 958–973 (Association for Computing Machinery, 2024).
Chan, A. et al. Infrastructure for AI agents. In Trans. Machine Learning Research (2025).
SAE International. Taxonomy and definitions for terms related to driving automation systems for on-road motor vehicles. SAE International https://www.sae.org/standards/content/j3016_202104 (2021).
Council of the European Union. Regulation (EU) 2024/1689 of the European Parliament and of the Council of 13 June 2024 laying down harmonised rules on artificial intelligence and amending Regulations (EC) No 300/2008, (EU) No 167/2013, (EU) No 168/2013, (EU) 2018/858, (EU) 2018/1139 and (EU) 2019/2144 and Directives 2014/90/EU, (EU) 2016/797 and (EU) 2020/1828 (Artificial Intelligence Act). Eur-LEX http://data.europa.eu/eli/reg/2024/1689/oj (2024).
Tabassi, E. Artificial intelligence risk management framework (AI RMF 1.0). National Institute of Standards and Technology http://nvlpubs.nist.gov/nistpubs/ai/NIST.AI.100-1.pdf (2023).
Goldman, A. I. A Theory of Human Action (Prentice Hall, 1970).
Davidson, D. I. in Agent, Action, and Reason (eds Binkley, R. W., Bronaugh, R. N. & Marras, A.) 1–37 (Univ. Toronto Press, 1971).
Dretske, F. Explaining Behavior: Reasons in a World of Causes (MIT Press, 1988).
Dung, L. Understanding artificial agency. Philos. Q. 75, 450–472 (2025).
Google Scholar
Ginet, C. On Action (Cambridge Univ. Press, 1990).
O’Connor, T. Persons and Causes: The Metaphysics of Free Will (Oxford Univ. Press, 2002).
Lowe, E. J. Personal Agency: The Metaphysics of Mind and Action (Oxford Univ. Press, 2008).
Frankfurt, H. G. Freedom of the will and the concept of a person. J. Philos. 68, 5–20 (1971).
Google Scholar
Taylor, C. in The Self: Psychological and Philosophical Issues (ed. Mischel, T.) 103–135 (Blackwell, 1977).
Grossman, S. J., & Hart, O. D. in Foundations of Insurance Economics (eds Dionne, G. & Harrington, S. E.) 302–340 (Springer, 1992).
Kauffman, S. & Clayton, P. On emergence, agency, and organization. Biol. Philos. 21, 501–521 (2006).
Google Scholar
Bouvard, V. et al. A review of human carcinogens—part B: biological agents. Lancet Oncol. 10, 321–322 (2009).
Google Scholar
Meincke, A. S. in Philosophy of Science (eds Christian, A., Hommen, D., Retzlaff, N. & Schurz, G.) 65–93 (Springer, 2018).
Damşa, C. I., Kirschner, P. A., Andriessen, J. E. B., Erkens, G. & Sins, P. H. M. Shared epistemic agency: an empirical study of an emergent construct. J. Learn. Sci. 19, 143–186 (2010).
Google Scholar
Elgin, C. Z. Epistemic agency. Theory Res. Educ. 11, 135–152 (2013).
Google Scholar
Sosa, E. Epistemic agency. J. Philos. 110, 585–605 (2013).
Google Scholar
Nickel, J. W. in Ethnicity and Group Rights (eds Shapiro, I. & Kymlicka, W.) 235–256 (New York Univ. Press, 1997).
List, C. & Pettit, P. Group Agency: The Possibility, Design, and Status of Corporate Agents (Oxford Univ. Press, 2011).
Tollefsen, D. Groups as Agents (Polity, 2015).
Bratman, M. Shared Agency: A Planning Theory of Acting Together (Oxford Univ. Press, 2014).
Shapiro, S. J. in Rational and Social Agency: The Philosophy of Michael Bratman (eds Vargas, M. & Yaffe, G.) 257–293 (Oxford Univ. Press, 2014).
Le Besnerais, A., Moore, J. W., Berberian, B. & Grynszpan, O. Sense of agency in joint action: a critical review of we-agency. Front. Psychol. 15, 1331084 (2024).
Google Scholar
Russell, S. J. Rationality and intelligence. Artif. Intell. 94, 57–77 (1997).
Google Scholar
Rao, A. S. & Wooldridge, M. in Foundations of Rational Agency (eds Wooldridge, M. & Rao, A.) 1–10 (Springer, 1999).
Silver, D. et al. Mastering the game of Go without human knowledge. Nature 550, 354–359 (2017).
Google Scholar
Brown, N. & Sandholm, T. Superhuman AI for multiplayer poker. Science 365, 885–890 (2019).
Google Scholar
Meta Fundamental AI Research Diplomacy Team (FAIR) et al. Human-level play in the game of Diplomacy by combining language models with strategic reasoning. Science 378, 1067–1074 (2022).
Google Scholar
Lake, B. M., Ullman, T. D., Tenenbaum, J. B. & Gershman, S. J. Building machines that learn and think like people. Behav. Brain Sci. 40, e253 (2017).
Google Scholar
Park, J. S. et al. Generative agents: interactive simulacra of human behavior. In UIST ’23: Proc. 36th Annual ACM Symposium on User Interface Software and Technology https://doi.org/10.1145/3586183.3606763 (ACM, 2023).
Wang, X. & Zhou, D. Chain-of-thought reasoning without prompting. Adv. Neural Inf. Process. Syst. 37, 66383–66409 (2024).
Google Scholar
Zelikman, E., Wu, Y., Mu, J. & Goodman, N. D. STaR: self-taught reasoner bootstrapping reasoning with reasoning. In Proc. 36th International Conference on Neural Information Processing Systems (NIPS’22) 15476–15488 (Curran Associates Inc., 2022).
Anthropic. Claude 3.5 Sonnet Model Card Addendum https://www-cdn.anthropic.com/fed9cc193a14b84131812372d8d5857f8f304c52/Model_Card_Claude_3_Addendum.pdf (2024).
Kavukcuoglu, K. Gemini 2.5: Our most intelligent AI model. Google https://blog.google/technology/google-deepmind/gemini-model-thinking-updates-march-2025/ (2025).
Jaech, A. et al. OpenAI o1 system card. Preprint at https://doi.org/10.48550/arXiv.2412.16720 (2024).
Berlin, I. in Four Essays On Liberty 118–172 (Oxford Univ. Press, 1969).
Dennett, D. C. Freedom Evolves (Viking, 2003).
Vagia, M., Transeth, A. A. & Fjerdingen, S. A. A literature review on the levels of automation during the years. What are the different taxonomies that have been proposed? Appl. Ergon. 53, 190–202 (2016).
Google Scholar
Klyubin, A. S., Polani, D. & Nehaniv, C. L. Empowerment: a universal agent-centric measure of control. In Proc. 2005 IEEE Congress on Evolutionary Computation 128–135 (IEEE, 2005).
Chalmers, D. J. Reality+: Virtual Worlds and the Problems of Philosophy (Allen Lane, 2022).
Spirtes, P., Glymour, C. N., & Scheines, R. Causation, Prediction, and Search (MIT Press, 2000).
Pearl, J. Causality: Models, Reasoning, and Inference 2nd edn (Cambridge Univ. Press, 2009).
Bareinboim, E., Correa, J. D., Ibeling, D., & Icard, T. in Probabilistic and Causal Inference: The Works of Judea Pearl (eds Geffner, H., Dechter, R. & Halpern, J. Y.) 507–556 (Association for Computing Machinery, 2022).
Jaeger, J., Riedl, A., Djedovic, A., Vervaeke, J. & Walsh, D. Naturalizing relevance realization: why agency and cognition are fundamentally not computational. Front. Psychol. 15, 1362658 (2024).
Google Scholar
Sacerdoti, E. D. Planning in a hierarchy of abstraction spaces. Artif. Intell. 5, 115–135 (1974).
Google Scholar
Georgievski, I. & Aiello, M. HTN planning: overview, comparison, and beyond. Artif. Intell. 222, 124–156 (2015).
Google Scholar
Kwa, T. et al. Measuring AI ability to complete long software tasks. Preprint at https://doi.org/10.48550/arXiv.2503.14499 (2026).
Deb, K., Sindhya, K. & Hakanen, J. in Decision Sciences: Theory and Practice (eds Sengupta, R., Gupta, A. & Dutta, J.) 145–184 (CRC Press, 2016).
Li, M. & Vitányi, P. An Introduction to Kolmogorov Complexity and Its Applications (Springer, 2019).
Shannon, C. E. A mathematical theory of communication. Bell Syst. Tech. J 27, 379–423 (1948).
Google Scholar
Papadimitriou, C. H. in Encyclopedia of Computer Science (eds Ralston, A., Reilly, E. D. & Hemmendinger, D.) 260–265 (Wiley, 2003).
Holland, J. H. in Complexity and Industrial Clusters (eds Curzio, A. Q. & Fortis, M.) 25–34 (Physica, 2002).
Legg, S. & Hutter, M. Universal intelligence: a definition of machine intelligence. Minds Mach. 17, 391–444 (2007).
Google Scholar
Goertzel, B. Artificial general intelligence: concept, state of the art, and future prospects. J. Artif. Gen. Intell. 5, 1–48 (2014).
Hutter, M. Universal Artificial Intelligence: Sequential Decisions Based on Algorithmic Probability (Springer, 2005).
Morris, M. R. et al. Position: Levels of AGI for operationalizing progress on the path to AGI. In Proc. 41st International Conference on Machine Learning (ICML’24) 36308–36321 (JMLR.org, 2024).
Ehtesham, A., Singh, A., Gupta, G. K., & Kumar, S. A survey of agent interoperability protocols: Model Context Protocol (MCP), Agent Communication Protocol (ACP), Agent-to-Agent Protocol (A2A), and Agent Network Protocol (ANP). Preprint at https://doi.org/10.48550/arXiv.2505.02279 (2025).
Anderljung, M. et al. Frontier AI regulation: managing emerging risks to public safety. Preprint at https://doi.org/10.48550/arXiv.2307.03718 (2023).
Kasirzadeh, A. Measurement Challenges in AI Catastrophic Risk Governance and Safety Frameworks (Tech Policy Press, 2024).
Luo, Z., Kasirzadeh, A., & Shah, N. B. The more you automate, the less you see: hidden pitfalls of AI scientist systems. Preprint at https://doi.org/10.48550/arXiv.2509.08713 (2025).
Bowman, S. R. et al. Measuring progress on scalable oversight for large language models. Preprint at https://doi.org/10.48550/arXiv.2211.03540 (2022).
Carpenter, J. in Living with Robots: Emerging Issues on the Psychological and Social Implications of Robotics (eds Park, R. de Visser, E. J. & Rovira, E.) 75–90 (Academic, 2020).
Bereska, L. & Gavves, E. Mechanistic interpretability for AI safety – a review. Preprint at https://doi.org/10.48550/arXiv.2404.14082 (2024).
Hacker, P., Edwards, L. & Kasirzadeh, A. AI, digital platforms, and the new systemic risk. In Proc. 2026 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’26) https://doi.org/10.1145/3805689.3806506 (ACM, 2026).
Frey, C. B. & Osborne, M. A. The future of employment: how susceptible are jobs to computerisation? Technol. Forecast. Soc. Change 114, 254–280 (2017).
Google Scholar
Webb, M. The impact of artificial intelligence on the labor market. SSRN Scholarly Paper at https://doi.org/10.2139/ssrn.3482150 (2020).
Tyler, S. et al. Use of artificial intelligence in triage in hospital emergency departments: a scoping review. Cureus 16, e59906 (2024).
Google Scholar
Salge, C., Glackin, C., & Polani, D. in Guided Self-Organization: Inception (ed. Prokopenko, M.) 67–114 (Springer, 2014).
Luccioni, S., Jernite, Y. & Strubell, E. In Proc. 2024 ACM Conference on Fairness, Accountability, and Transparency (FAccT ’24) 85–99 (Association for Computing Machinery, 2024).
Franklin, S. & Graesser, A. in Intelligent Agents III Agent Theories, Architectures, and Languages (eds Müller, J. P., Wooldridge, M. J. & Jennings, N. R.) 21–35 (Springer, 1997).
Shavit, Y., et al. Practices for governing agentic AI systems. Research Paper, OpenAI https://cdn.openai.com/papers/practices-for-governing-agentic-ai-systems.pdf (2023).
Masterman, T., Besen, S., Sawtell, M. & Chao, A. The landscape of emerging AI agent architectures for reasoning, planning, and tool calling: a survey. Preprint at https://doi.org/10.48550/arXiv.2404.11584 (2024).
Huang, X. et al. Understanding the planning of LLM agents: a survey. Preprint at https://doi.org/10.48550/arXiv.2402.02716 (2024).
Li, J. et al. EvoCodeBench: an evolving code generation benchmark with domain-specific evaluations. Adv. Neural Inf. Process. Syst. 37, 57619–57641 (2024).
Google Scholar
Sharkey, L. et al. Open problems in mechanistic interpretability. Preprint at https://doi.org/10.48550/arXiv.2501.16496 (2025).
For more tech updates, stay tuned to our blog.

















