Summary
This paper is interested in the use of LLMs to assist humans in critical thinking. Academic philosophers, in interviews which are qualitatively coded, are asked to reflect on the use of LLMs to assist them in their work. Overall, the interviewed philosophers find that LLMs are not useful. The paper diagnoses this uselessness in two missing properties in LLMs: they lack a sense of selfhood and initiative. Possible roles for LLMs that vary these features (i.e., having selfhood and lacking initiative, lacking selfhood and having initiative, and having both) are discussed as a way of trying to guide the development of LLMs.
Reasons to accept
I found the discussion of the roles for LLMs (interlocuter, monitor, and respondent) to be interesting. I think this framing is interesting to consider in an educational context (e.g., how should universities position themselves with relation to AI assistants). Further the paper is well written and does a solid job of situating the current use cases of LLMs in contrast with their use in fostering critical thinking.
Reasons to reject
The link between the paper’s aim (positioning LLMs as critical thinking tools) and the argumentation and methods (the interview of academic philosophers) is not clear. If the aim is to study what makes a tool useful for supporting critical thinking, why are philosopher’s views on LMs (e.g., answers to questions like “What are some risks and weaknesses for language models in philosophy”) relevant for addressing this? At times the paper discusses how critical thinking is deployed in endeavors like “history, science, and philosophy”, wouldn’t it make sense to probe how each of these fields utilizes critical thinking and what tools are useful for doing this more efficiently?
Further, the discussion of the things that are useful for philosophers in critical thinking, selfhood and initiative, both require agency (in my understanding) in the definitions given in the paper (“selfhood is a resource’s ability to have certain locally persistent internal states (such as perspectives, beliefs, opinions, memory) and to consistently use them as the basis for judgements” and “Initiative is a resource’s ability to set its own intentions and goals, possibly different from its user’s, and to execute actions oriented towards those intentions.”). This feels not concrete enough to offer practical guidance in developing LLMs.
Questions to authors
Who is the “we” in “we claim all the rights to think?”. I read this as saying we humans claim all the rights to think. What relationship does that have to LLMs? It feels to me, in fact, like that statement is in tension with the aims of the paper (the description of a tool to automate some types of thinking). Perhaps, we will use LLMs to think more critically? But the other uses cases (e.g., coding, generating emails) mentioned in the beginning are taken as useful because they alleviate the need for that type of thinking.
Two issues with LLMs, that they are “highly neutral, detached, and non-judgmental” and “servile, passive, and incurious” are mentioned a few times in the paper, however, no empirical results are given to support this. If this is the perception of the interviewees that is fine, but it should be made clear in the paper.