







ABSTRACT The success of large language models (LLMs) across many domains of AI research has generated intense debate. Some attribute their impressive performance on complex tasks to human‐like linguistic and cognitive capacities, whereas others ascribe it to shallow pattern matching. These disputes stem from deep‐seated philosophical disagreements about the nature of language and cognition. We provide an opinionated survey of these disagreements across core topics in the philosophy of mind and language, including syntactic competence, compositionality, linguistic meaning, representation, attitudes, reasoning, agency, and consciousness. We contend that progress on these issues requires not only clarity about background philosophical commitments but also, in many cases, close engagement with emerging empirical evidence.
Epistemological Fault Lines Between Human and Artificial Intelligence
Large language models (LLMs) are widely described as artificial intelligence, yet their epistemic profile diverges sharply from human cognition. Here we show that the apparent alignment between...

A Rational Analysis of the Effects of Sycophantic AI
People increasingly use large language models (LLMs) to explore ideas, gather information, and make sense of the world. In these interactions, they encounter agents that are overly agreeable. We...

We Don't Understand Neural Networks At The Algorithmic Level
The largest ongoing debate about AI is “Are Large Language Models (LLMs) intelligent?” That makes sense, at least: the evidence is ambiguous and the stakes a...
Sense-making reconsidered: large language models and the blind spot of embodied cognition
Large Language Models (LLMs) demonstrate a kind of linguistic competence that theories of embodied and enactive cognition have long deemed impossible for systems lacking the meaningful perspective of a living being, i.e., the capacity for sense-making. Facing up to this unexpected technological development requires confronting what I propose to call the “AI dilemma”: either frontier LLMs are capable of sense-making despite lacking biological embodiment, or the kind of linguistic competence they exhibit does not necessarily require sense-making. In their chapter on cognition, Frank, Thompson, and Gleiser (2024) maintain that no AI system comes close to realizing relevance, a position that derives much of its motivation from past practical failures. However, frontier LLMs have effectively overcome Dreyfus’ commonsense knowledge problem, such that their dismissal as categorically mindless risks undermining Frank et al.’s central claim that human cognition is deeply intertwined with lived experience. I therefore argue in favor of the alternative side of the AI dilemma: human-level linguistic competence of LLMs should be recognized as a novel non‑biological form of sense‑making, based on a technologically‑mediated embodiment whose enabling properties are in need of further theoretical analysis. This reorientation invites enactive theory to clarify which aspects of sense-making may be universal and which aspects are specifically contingent on organic life, thereby advancing its conceptual framework in dialogue with contemporary AI.
Understanding Understanding: A Pragmatic Framework Motivated by...
Motivated by the rapid ascent of Large Language Models (LLMs) and debates about the extent to which they possess human-level qualities, we propose a framework for testing whether any agent (be it...

The Thoughts The Civilized Keep
The hype around a new AI language generator reveals the sterility of mainstream thinking on AI today — and indeed on how we think about thinking itself.

Model Collapse Ends AI Hype
Large language models are not the problem
If a Large Language Model (LLM) can replicate your scientific contribution, the problem is not the LLM. What does it say about our field that so much of the anxiety about AI comes down to the fear that a machine could do what we do? Perhaps it says we should be doing something better.

Can generative artificial intelligence be considered a cognitive subject? An analytic analysis
This paper examines whether contemporary generative artificial intelligence (GAI), especially large language models (LLMs), can be regarded as a “cognitive subject” in the epistemic sense relevant to the production and endorsement of knowledge claims. GAI systems increasingly participate in writing, research, and decision-making workflows and can display striking competence in information processing and task-directed problem solving. Yet, the thesis that GAI is a cognitive subject is stronger than the observation that GAI contributes as a cognitive tool. Therefore, we propose an explicit set of necessary and sufficient conditions for cognitive subjecthood and evaluate each condition in light of recent philosophical and empirical scholarship. The analysis supports a two-part conclusion: (i) present-day GAI can reasonably be described as a cognitively significant contributor to knowledge production, but (ii) it does not satisfy the conditions for cognitive subjecthood, largely because robust intentionality, metacognitive self-representation, and consciousness-related indicator properties are not established.
Big AI is accelerating the metacrisis: What can we do?
The world is in the grip of ecological, meaning, and language crises that are converging into a metacrisis. Big AI is accelerating them all. LLM engineering sits at the core. Despite the public good motives of language engineers and the promise of LLMs, this work is being leveraged to create unprecedented wealth and power for a handful of individuals and corporations while causing existential harm to life on earth. As a profession, we urgently need to come together to explore alternatives and to design a life-affirming future for our field of natural language processing that is centered on human flourishing on a living planet.

Small Language Models are the Future of Agentic AI
Large language models (LLMs) are often praised for exhibiting near-human performance on a wide range of tasks and valued for their ability to hold a general conversation. The rise of agentic AI systems is, however, ushering in a mass of applications in which language models perform a small number of specialized tasks repetitively and with little variation. Here we lay out the position that small language models (SLMs) are sufficiently powerful, inherently more suitable, and necessarily more economical for many invocations in agentic systems, and are therefore the future of agentic AI. Our argumentation is grounded in the current level of capabilities exhibited by SLMs, the common architectures of agentic systems, and the economy of LM deployment. We further argue that in situations where general-purpose conversational abilities are essential, heterogeneous agentic systems (i.e., agents invoking multiple different models) are the natural choice. We discuss the potential barriers for the adoption of SLMs in agentic systems and outline a general LLM-to-SLM agent conversion algorithm. Our position, formulated as a value statement, highlights the significance of the operational and economic impact even a partial shift from LLMs to SLMs is to have on the AI agent industry. We aim to stimulate the discussion on the effective use of AI resources and hope to advance the efforts to lower the costs of AI of the present day. Calling for both contributions to and critique of our position, we commit to publishing all such correspondence at https://research.nvidia.com/labs/lpr/slm-agents.

We Need to Talk About How We Talk About 'AI'
We share a responsibility to create and use empowering metaphors rather than misleading language, write Emily M. Bender and Nanna Inie.

AI learns language from skewed sources. That could change how we humans speak – and think | Bruce Schneier
Large language models aren’t trained on real-life conversations. As we encounter their language, it could affect our own

Cognitive Architectures for Language Agents
Recent efforts have augmented large language models (LLMs) with external resources (e.g., the Internet) or internal control flows (e.g., prompt chaining) for tasks requiring grounding or reasoning, leading to a new class of language agents. While these agents have achieved substantial empirical success, we lack a framework to organize existing agents and plan future developments. In this paper, we draw on the rich history of cognitive science and symbolic artificial intelligence to propose Cognitive Architectures for Language Agents (CoALA). CoALA describes a language agent with modular memory components, a structured action space to interact with internal memory and external environments, and a generalized decision-making process to choose actions. We use CoALA to retrospectively survey and organize a large body of recent work, and prospectively identify actionable directions towards more capable agents. Taken together, CoALA contextualizes today’s language agents within the broader history of AI and outlines a path towards language-based general intelligence.
