# Generating Philosophy with AI 28 Apr 2026 > [!NOTE] > this is LLM text that I have shouted into existence. currently rewriting it line by line. ## Introduction The last decade or so has seen the rise of generative artificial intelligence: systems that produce text, images, code, music, video, and other outputs in response to prompts. AI has had success in domains where the value of an output is not exhausted by its superficial fluency. For example, in February 2026, researchers working on gluon scattering amplitudes gave GPT-5.2 worked examples for three, four, five, and six particles and asked it to find the general formula. GPT-5.2 proposed a closed-form expression; another internal model supplied a proof; and the authors then verified the result. The resulting paper argues that single-minus tree-level gluon amplitudes, often presumed to vanish, are non-vanishing in certain half-collinear configurations (Guevara et al. 2026). There are also recent examples in mathematics (Novikov et al. 2025), biomedicine (Gottweis et al. 2025), and materials science (Zeni et al. 2025). In this paper we argue that we should expect similar success in philosophy. Specifically, we argue that current-generation LLMs are capable of producing philosophical texts that are _worth reading_. This phrase might seem loose, but that is part of its point. We do not want to begin by settling what counts as _good_ philosophy. Instead, we appeal to a distinction that anyone reading this text will recognise. You have read texts that are worth reading, and you have read texts that are not. As you begin reading this article, you likely hope that it is worth reading, in the sense that the time spent reading it will not be wasted. When you write a philosophical text yourself you aim to make it worth readers' while to read it, and whether or not the journal you send it to accepts it, depends on whether or not they agree. Two clarifications are needed. First, a text’s being worth reading is not the same as its being correct. A text can repay attention even if one rejects its conclusion: it may sharpen a distinction or answer an objection in a way that changes the dialectical situation. Second, the minimal unit we are concerned with is not the bare conclusion of an argument, but the argument itself. If an LLM output consists only in a pronouncement on some philosophical topic ('Direct Realism is correct', 'We should be utilitarians'), it is hard to see why it would be worth reading in and of itself, for the same reason that a bare pronouncement by a human philosopher would not be worth reading.[^1] The next three sections develop the main argument. Section I rejects the challenge from authorship: the claim that an LLM output cannot be philosophy worth reading because no philosopher lies behind it. Section II turns to abduction and argues that the absence of human-style inference to the best explanation in the producer does not preclude abductive structure in the product. Section III considers phenomenology and argues that the lack of consciousness does not prevent LLMs from producing philosophy grounded in phenomenology. ## I. The challenge from authorship In this section we address what we might call the _challenge from authorship_: the idea that philosophy is something that only persons, or at least minds, can do. This view has not, to our knowledge, been explicitly defended in just this form, but it gives shape to an intuition that many philosophers may have: philosophy is a person-only domain. An imperfect comparison is with art. One might deny that an image generated by an AI system, at least in the familiar prompt-and-output cases, is an artwork because no artist exercises the relevant kind of intentional control over its production. One might think, for similar reasons, that philosophy can only be done by people. No text produced by an LLM can be a work of philosophy, because no philosopher lies behind it. One might then think, for parallel reasons, that philosophy too requires a philosopher: no text produced by an LLM can be a work of philosophy, because no philosophical agent lies behind it. Consider also that, like art, the study of philosophy is often focussed on individuals. Philosophy undergraduates take a course on Kant's ethics, or Lewis' metaphysics, and even at more advanced levels one finds specialists, conferences etc. spotlighting the work of specific philosophers. Compare this to the sciences: as a rule, scientific ideas, theories, discoveries etc. are the focus, not the individuals behind them: one does not find scientists who specialise in the work of Newton, or of Einstein; nor do biology departments teach Crick's view of DNA rather than Watson's. %%is this paragraph accurate re: science?%% We will try now and make this intuition more precise by continuing the comparison with artworks and philosophical works. We shall do this by considering the degree to which Davies' *performance* theory of art can be transposed to philosophy. He writes: > The work — what the artist achieves — is the process eventuating in that product. Works themselves are neither structures nor objects simpliciter, nor are they contextualized structures or objects [...]They are, rather, intentionally guided generative performances that eventuate in contextualized structures or objects. (p. 98) On Davies’ view, when a painter paints a picture, the canvas is what we attend to, but it is not the work. The work is the artist’s intentionally guided activity in producing that canvas; the canvas is, in Davies’ terms, the "focus of our appreciative interest in the work" (2004, p. 151). This is why provenance matters to him in a deeper way than it would matter on a view that identifies the artwork with a product plus contextual properties. Facts about how the object came into being help determine what the work is and what is properly appreciated in it. If the same model were transposed to philosophy, an LLM text would fail not because it is badly argued, but because the relevant kind of philosophical performance is missing. Consider what is involved in attending to a Vermeer. We are not only registering a perceptual surface%%not how i write%%. We are taking that surface as the outcome of a certain painter’s activity, in a certain historical context, with certain resources and limitations. Davies presses this point through cases in which perceptual sameness, or near-sameness, fails to settle artistic identity or appreciation.%%not how i write%% A canvas might emerge by accident from a washing machine and happen to look like a Rembrandt%%is this a danto example? look it up. if it is it needs to be referenced.%%; in that case, there is a Rembrandt-like surface, but no artistic performance of the relevant kind. Or a canvas might be presented as a Vermeer when it was in fact painted by van Meegeren%%famous forger? if so, you need to mention, even if just in a footnote%%; in that case, there is an artistic performance, but not the one the work was taken to make available. The point is not just that provenance gives us extra information. It is that provenance can change what we take the work to be and what kind of achievement we take ourselves to be appreciating. If Davies is right, the surface does not by itself settle the work. Here is the analogous proposal for philosophy. A philosophical text is not itself the philosophical work. The text is the product of the thinking, writing, and philosophising done by a person or group of persons over time. The text is therefore the focus of our attention, but only as a way of accessing the philosophical performance that brought it into being. On this proposal, a philosophical work is not identical with the sequence of sentences on the page. The text is the product of someone’s activity of thinking through a problem and giving that activity argumentative form. Reading the text is then a way of engaging with that activity: not just with a conclusion, but with the route by which the conclusion is reached.%%not how i write%% The authorship challenge is therefore not just a worry about missing biography. It is the stronger claim that, if no one has done the relevant philosophising, there is no philosophical work to which the text gives access. The question is whether this transposition should be accepted. We do not think it should. Davies has a reason to move from product to performance in the case of art: production history can affect which work we are dealing with and what is available for appreciation. A Rembrandt-like surface produced by accident is not a Rembrandt; a van Meegeren presented as a Vermeer is not the work it is taken to be. The philosophical case is different. If two texts contain the same argument, including the same inferential moves, the same considerations count for and against them. Their philosophical merit does not vary with the route by which the words came to be written. When we assess a philosophical paper, we ask whether the text does philosophical work. Does it introduce a distinction that helps? Does it answer an objection that would otherwise remain pressing? These questions do not require us to look behind the text to the philosopher’s activity. The grounds for the judgement lie in the argument as presented, not in the history of its production. This is where the analogy with Davies breaks down. Two papers that read identically do not differ in argumentative merit: they make the same moves and face the same objections. In the art case, production history can change what the work is. In the philosophy case, it changes, at most, what we think about the producer or the process by which the text came about. The organisation of analytic philosophy reflects this. Journals often strip author information from submissions before sending them to referees, and they do so because facts about authorship are treated as possible sources of distortion. The point is not that blind review always succeeds, or that philosophical practice is never interested in authors. The point is narrower: in this central evaluative context, the paper is supposed to be assessed by attending to what it says, not by reconstructing the circumstances under which it was written. A point from Dellsén et al. (2024) helps to articulate the same thought, although their concern is philosophical progress rather than LLM authorship. On their view, philosophical progress is "for-whom" rather than "by-whom": it consists in putting people in a position to increase their understanding, usually by making philosophical ideas publicly available (2024, p. 679). For present purposes, the useful thought is that philosophy makes its contribution through public materials that others can take up: arguments, theories, distinctions, thought experiments, and ways of framing problems. If this is correct, we should be cautious about locating the philosophical work behind the public text, in the process by which the text came about. The public text is not a dispensable trace of philosophy; it is where the philosophical contribution becomes available. The challenge from authorship is therefore a constitutive challenge. It treats the philosopher’s activity not merely as something that causes a philosophical work to exist, but as part of what the work is. On this picture, even a text indiscernible from a philosophical paper would not be philosophy if no philosophical activity lay behind it. We have argued that this should be rejected. If a novel philosophical text were produced by the wind blowing sand into a readable pattern, or by a very faulty washing machine, that would not, in and of itself, prevent the resulting text from being worth reading. What remain are not objections about what philosophy is, but objections about whether LLMs can produce texts with the relevant philosophical properties. That is, While the authorship challenge argued that text produced by an LLM cannot be philosophy worth reading in virtue of the fact that it was produced by an LLM, these *causal* challenges, on the other hand, can be thought of as claims that LLMs in their current state cannot produce philosophy worth reading, because LLMs lack certain capacities that are required to write worthwhile philosophy, or worthwhile philosophy in certain areas. In the next section, we consider the challenge from abduction, %%extremely succinct description of next section%%Section 3 turns to the parallel concern that some philosophical texts require phenomenal materials available only to conscious subjects. ### Footnotes 1. reference the ai image literature here, and mention that in most cases it is hard to imagine images being created without a human influencing things at least in some way. [↩](#user-content-fnref-1) 2. Note that such a view does not amount to the denial that LLMs can produce beautiful images. We shall return to this point later. [↩](#user-content-fnref-2) ## II. The challenge from abduction Much philosophical theorising proceeds by inference to the best explanation. A philosopher offers an account of some phenomenon and defends it by arguing that, if true, it would explain the relevant evidence better than its rivals. Williamson treats this as a legitimate method of argument in philosophy: philosophy, on this view, often advances by comparing theories with respect to their explanatory power, their fit with the evidence, and their theoretical virtues (Williamson 2016, pp. 351–356). The challenge is straightforward. If LLMs do not perform inference to the best explanation, it may seem that they cannot produce philosophical texts whose value depends on abductive argument. Floridi et al. give this challenge a precise form. They write: LLMs seem to perform a kind of zeroth-order abduction: given a prompt, they generate a plausible continuation (a hypothesis or explanation) based purely on learned associations. In reality, their operation is driven by maximising the probability of the sequence... The model does not understand what an explanation is, but it produces text that follows the typical phrasing and structure of explanations. It does not reason about causes from scratch but outputs typical causes for typical effects observed in the training data. (Floridi et al. 2025, p. 9) The claim is not that LLMs cannot produce text that looks explanatory. They often can. The claim is that such text is generated by learned associations and sequence probability, not by an understanding of evidence, causes, truth, or explanation. What appears to be abductive reasoning is, on their view, the surface result of a stochastic process. Floridi et al. are right about the process. An LLM does not understand a phenomenon as calling for explanation. It does not knowingly generate live candidate explanations, compare them, and infer the one that would best explain the data. It has no grasp of one candidate as lovelier or likelier than another. We should not respond by saying that LLMs secretly perform human-style inference to the best explanation. The question is instead whether a text produced by such a system can contain a good abductive argument. To see why it can, recall what Lipton’s account of inference to the best explanation assesses. On his view, we infer "what would, if true, provide the best explanation" of the evidence (Lipton 2004, p. 56). The phrase ‘if true’ is doing real work. We do not first identify the actual explanation and then infer it; that would require us to have reached the end of inquiry before inquiry begins. We assess potential explanations: candidates that would explain the data if they were true (Lipton 2004, pp. 57–59). A potential explanation is the sort of thing that prose can present. A text can specify the data, formulate the candidate, identify the relevant contrast, compare live alternatives, and show what the candidate would explain if true. This is where Lipton’s distinction between the likeliest and the loveliest explanation matters. The likeliest explanation is the one most likely to be true; the loveliest explanation is the one that would provide the most understanding if it were true. As Lipton puts it, "Likeliness speaks of truth; loveliness of potential understanding" (2004, p. 59). If inference to the best explanation meant only inference to the likeliest candidate, the account would say little more than that we infer what we judge most probable. Lipton’s stronger claim is that explanatory virtues help guide judgments of likelihood: loveliness is, at least sometimes, a guide to likeliness (2004, pp. 60–62). Williamson gives the corresponding point in philosophical terms when he says that a theory should be unified, not arbitrary, gerrymandered, ad hoc, or messily complicated; in short, it should combine simplicity with strength (Williamson 2016, p. 354). These are features of theories as they are articulated. They are visible in the text. Lipton also shows that abductive reasoning does not begin from the whole space of logical possibilities. Inquiry normally starts from a restricted set of live candidates. We first identify serious candidates, then compare them (Lipton 2004, p. 59). This matters because the first filter is itself part of philosophical practice. Philosophers inherit a structured background of distinctions, problems, objections, examples, and candidate views. That background shapes what counts as a live option in the first place. A paper that proposes a theory of perception, depiction, consciousness, or reference does not compare it with every logically possible alternative. It situates it within a debate whose options have already been shaped by previous argument. The philosophical corpus is one such background. It is not a neutral heap of sentences about philosophical topics. It is the written record of claims, objections, distinctions, revisions, and failed proposals that have been taken up and tested within philosophical practice. This does not mean that everything in the corpus is good philosophy, or that what survives is true. It means that the corpus is partly structured by past philosophical selection. Arguments are repeated because they are useful; distinctions persist because they do work; objections are preserved because they expose pressure points. The corpus therefore contains not only philosophical vocabulary, but traces of the abductive and dialectical standards by which philosophical texts have been produced and assessed. This gives us the mechanism. An LLM does not cease to be a next-token predictor when it produces philosophy. It samples a token from a learned conditional distribution, appends that token to the context, and repeats the process. But the distribution from which it samples has been trained on texts in which philosophical patterns are already present. When the training corpus contains abductively structured philosophical writing, the model’s conditional probabilities are shaped by that structure. The model is not judging that a candidate explanation is better than its rivals. Rather, it is generating a trajectory through a space of possible continuations whose local probabilities have been shaped by earlier philosophical texts. The terminology of semiotic physics is useful here, provided it is used sparingly. A generated text is a trajectory: the prompt plus the output-so-far after each step of the autoregressive loop. The model supplies transition probabilities over possible next tokens; sampling and appending a token produces the next state; repeated application produces the full continuation (Jan 2023; metasemi 2023). The heavier parts of the framework are not needed for the present argument. What matters is the local-to-global point. A philosophical argument is not a single token, but an extended trajectory. If the local transition tendencies have been shaped by a corpus in which abductive structures are common, then the resulting trajectory can display abductive structure at the level of the argument. Floridi et al. themselves say that LLMs have "absorbed patterns of human abductive reasoning as expressed in writing" (2025, p. 9). That sentence should not be inflated into the claim that LLMs understand abductive reasoning. But it should not be deflated into the claim that they have acquired only empty verbal templates. If abductive reasoning is expressed in writing, and if philosophical writing is one of the places where such reasoning is refined, criticised, and transmitted, then training on philosophical writing can shape the model’s generative tendencies in abductively relevant ways. The model does not need to perform the earlier reasoning in order for its outputs to bear the public traces of that reasoning. The result is a product-side capacity. A text produced by an LLM can formulate a potential explanation, place it against live alternatives, and display virtues relevant to abductive assessment. It can show why one distinction handles a case better than another, why an objection presses on a theory, or why a debate has been framed around the wrong contrast. None of this entails that the text is correct. It also does not entail that the model understood what it was doing. But it does mean that the text can make available an object of philosophical assessment: a candidate explanation whose merits can be judged by readers. The challenge from abduction therefore does not show that LLM-generated philosophy is impossible. It shows that the relevant capacity cannot be located in a human-like act of abductive judgement by the model. That concession is harmless if the claim concerns the product rather than the producer. LLMs do not perform inference to the best explanation in the way philosophers do. Still, given a philosophical corpus shaped by past abductive selection, they can produce texts that contain potential explanations, organise live alternatives, and exhibit explanatory virtues. Whether a particular output succeeds is then assessed in the ordinary philosophical way: by reading the text and asking what, if anything, it explains. ## III. The challenge from phenomenology ### Experiential Axioms A further capacity worry concerns phenomenology. Few would say that LLMs are conscious, and we will assume the same here; yet this might seem to pose a problem for LLM philosophy, or at least for philosophy grounded in, or making use of, phenomenology. Some philosophy interrogates or refers to what it is like to see red (Harman, 1990), to feel anger (Goldie, 2000), or to have a particular intuition take hold (Chudnoff, 2011). If LLMs lack conscious experience, it seems as if this might hamper their ability to produce worthwhile philosophy which relies on it. This is not to say that all philosophy would be off bounds: large stretches of philosophy of language and modal metaphysics proceed without leaning on the phenomenology of any particular experience. Zahavy’s discussion of a thought experiment of Einstein's brings out this worry: > Einstein’s variation required inventing new axioms based on a physical intuition that did not yet exist in the mathematics. He envisioned a physicist inside an elevator being uniformly accelerated through deep space. Inside this enclosure, the sensory experience reveals a specific pattern: when objects are released, the floor rushes up to meet them. To the physicist, the objects appear to fall with identical acceleration, regardless of composition. Thus, the simulation here was not a permutation of symbols, but a manipulation of perceptual experience. (Zahavy 2026, §5) The thinker imagines[^4] some set of circumstances and attends to what would be experienced within it — in Einstein’s case, that all objects inside the elevator would appear to fall with identical acceleration. That observation becomes the new axiom: a starting point arrived at through experiential simulation rather than formal derivation, from which further reasoning proceeds. If thinking of this kind depends on simulated experience, then it would seem to be out of reach for LLMs. They can provide descriptions of weightlessness or elevators, but they have never felt the sensation of an elevator descending, let alone weightlessness.[^2] Philosophy also uses experience based thought experiments. Jackson’s Mary case turns on what it is like to see colour, and we might think that as with Einstein's thought experiment, it provides us with an experiential axiom, from which further philosophical reasoning can proceed. The same worry then arises in philosophy: experience based thought experiments seem to require what LLMs do not have.[^3] ### Articulated Phenomenology Pigliucci offers an account of philosophy on which it is constrained by, but does not aim at, the world as the natural sciences do. He writes: > This means that the basic parameters that philosophers use as their inputs, the starting points of their philosophizing, their equivalent of axioms in mathematics and assumptions in logic (or rules in chess) are empirical data about the world. This data comes from both everyday experience [...] and of course increasingly from the world of science itself. (Pigliucci, p. 6) Philosophy begins from worldly materials, but those materials function as starting points for conceptual exploration. They are not used in the same way a physical datum is used to confirm or disconfirm an empirical theory.%%one more sentence, a good one, will make this paragraph substantial%% Pigliucci elaborates this picture by drawing on Smolin’s account of evocation, taking chess as the paradigm. Positing the rules of a game does not require that they pre-exist; once posited, they generate a structure with rigid properties — a space of consequences that can be explored but not chosen. Once the rules of chess are codified, all the facts about chess become demonstrable, even though chess did not exist before its rules were written down. Pigliucci’s claim is that philosophy operates in this register. Unlike the rules of chess or the axioms of mathematics, however, the starting points of philosophising are constrained empirically. They are constrained by how the world actually is, including by what experience is like. This is why Pigliucci distinguishes philosophy from fiction: philosophy is not merely the invention of imaginary possibilities. As he puts it: > Philosophy, I maintain, is in the business of doing empirically informed evoking, not inventing. (Pigliucci, p. 7) The same picture covers philosophical thought experiments. Even when philosophers explore possible worlds or imagined scenarios, they do so “with an interest in figuring things out as far as this world is concerned” (Pigliucci, p. 7). The thought experiment articulates an axiom — an experiential or empirical starting point — and the philosophical work proceeds within the conceptual landscape that axiom evokes. This brings out a difference between the elevator and Mary cases. Both are evocations of the kind Pigliucci describes: each posits an experiential axiom and develops what follows from it. What differs is what the evocation is for. In Einstein’s case, the evoked structure yields a hypothesis whose status is then settled by experiment — the elevator gave him the equivalence principle, but the principle’s truth was a matter for empirical confirmation. In Mary’s case, the evoked landscape is itself the object of inquiry; the philosophical question is what the landscape contains, not whether anything outside it corresponds. The role of the evocation, not its presence, is what tracks the disciplinary difference. Evocation is present in both cases; what differs is whether the evoked structure is the means to an external test or is itself the object of inquiry. No competent discussant of the knowledge argument has personally undergone her transition. Once the case is articulated, work on it is work on the articulation. Responses to Jackson press at the level of the articulated structure, not at the level of any discussant’s experience. Lewis’s reply, for instance, modifies what is taken to follow from Mary’s situation, not what Mary’s situation is taken to be like from the inside. What allows the Mary case to do philosophical work in public is its articulation: the experiential material it draws on has been made available in language. This is the form in which phenomenology enters philosophy generally. The articulation is what does the philosophical work; the experience the articulation refers to need not be undergone by the people working on it. Philosophers work on the experiences of the blind and on the experiences of non-human animals without first-hand access to either, by working on the articulations the literature has accumulated. The point matters for LLMs in a particular way. They have no raw phenomenology of their own; but no text corpus contains raw phenomenology either. What a corpus contains is articulated phenomenology, and it is in articulated form that phenomenology becomes usable in philosophical argument. Merleau-Ponty’s discussion of self-touch raises a sharper question — that of phenomenological _discovery_. Suppose the toucher-touched asymmetry was first identified by Merleau-Ponty himself, by sustained attention to his own embodied experience. The asymmetry would then be a phenomenological axiom out of reach of any LLM not trained on Merleau-Ponty or his interlocutors: an axiom an LLM could not have produced for itself, because the system lacks the body and the experience that the discovery requires. When one fingertip touches another, one finger plays the role of toucher and the other of touched. The roles can reverse, but not simultaneously: at any given instant, the body is split between touching and touched. But that does not prevent an LLM from working philosophically on the description once articulated. What survives, then, is a narrower asymmetry. Even granting that LLMs can work within articulated landscapes, some phenomenological articulations seem to be originated through first-person attention; LLMs have no experience to attend to. First-person attention is one route to an articulation; it is not what gives an articulation philosophical use. What makes an articulation philosophically usable, on Pigliucci’s picture, is not its causal origin but its functioning as an axiom — its capacity to evoke a landscape with rigid properties. An articulation can also be arrived at by working from the articulations a corpus already contains, generating new ones by extension and recombination. Whether a candidate articulation succeeds is a question about what it evokes, and that question is answered the way other philosophical questions are — by the public assessment of the conceptual structure the articulation makes available. It is the assessment any candidate articulation, whatever its origin, must finally meet. The phenomenology objection rests on a producer-to-product inference: that the absence of experience in the producer must remove phenomenological value from the product. The inference fails. LLMs lack conscious experience, but phenomenology enters philosophy as articulated content. Pigliucci’s account explains why this is not a workaround. Philosophy uses empirical and experiential materials by turning them into constrained spaces for conceptual exploration. Since those spaces are public and inferentially usable once articulated, current models can produce phenomenology-based philosophy worth reading. [^2]: footnote saying that he calls it manipulative abduction. it should probably also explain why we might think go this as abduction as well as what we talked about in the previous section [^3]: A nice example in the footnote will be the feeling of understanding that is sometimes used as a way of motivating cognitive phenomenology. ## IV. Authorship redux: the challenge from elicitation If current LLMs can produce philosophy worth reading, why are we not surrounded by great LLM philosophical texts? The ordinary experience of using these systems seems to support scepticism. Asked for philosophy, they often produce competent but lifeless exposition—paragraphs that tour a topic without ever applying pressure to it. A vague topic prompt does not ask for a philosophical intervention. It asks for the most probable kind of text under that topic-label, and in the training distribution that is often hedged survey, because the bulk of philosophical text written at that level of generality takes that form. The prompt is generic, and so is the region of the model’s learned space it activates. If interesting LLM philosophy appears only under careful prompting, perhaps the LLM is not really producing the philosophy after all. Perhaps the human prompter is producing philosophy by using the LLM. This version of the worry turns on control. The final text may be worth reading, but if the human fixes the task and chooses the result, the value can seem to belong to the human-guided process rather than to the model’s production. So “produced by an LLM” has to mean more than “appearing in an LLM output window”. A model may output philosophy worth reading without producing the features that make it worth reading. Take the grammar-correction case. A philosopher writes a brilliant argument and asks an LLM only to correct its punctuation. The LLM’s response may contain philosophy worth reading, but the model has not produced the argument in virtue of which the text is worth reading. It has output worthwhile philosophy without producing what makes it worth reading. So there is a continuum of contribution. At one end, the model merely polishes a human-produced argument; at the other, the human specifies a task and the model generates the philosophical move that makes the output worth reading. The cases that support the present thesis lie towards the latter end. Where the human supplies the argument and the model improves the prose, the LLM has not produced philosophy worth reading. Where the human specifies a problem and the model supplies the objection or distinction that makes the text worth reading, the output is LLM-produced in the sense that matters here. The elicitation objection assumes too simple a contrast between autonomous producer and mere tool. Elsewhere I have argued, with Terrone, that generative AI systems of the Midjourney type are best understood neither as agents nor as ordinary tools, but as generative systems with which users interact under conditions of partial control. LLMs are not intentional agents, but the ordinary tool model is also too crude. Prompting is not command execution. The prompt is not a blueprint that fixes the product in advance. It sets conditions under which the model generates. The user can constrain and iterate, but cannot determine every relevant feature of what emerges. Prompting is therefore a form of elicitation. A call for papers elicits philosophical answers without authoring them; an interlocutor in conversation elicits arguments from another philosopher without coming to be their author. So a prompt’s having elicited an LLM output does not by itself show that the prompter has supplied the philosophical content that makes the output worth reading. Selection is not generation either. A journal selects the papers it publishes but does not thereby produce them, and a reader’s selection of a good LLM output is, similarly, an act of assessment rather than of production. If philosophical corpora encode abductive and dialectical structure in the way I argued earlier, skilled prompting should aim to elicit those structures rather than request prose about a topic. Pigliucci’s account gives this a further formulation. A good philosophical prompt evokes a constrained conceptual space: it makes a problem determinate by placing a contrast under dialectical pressure, but leaves the philosophical move for the model to make. The point is not to insert phrases such as “one might object,” as though they were magic words. Such phrases matter only because they mark argumentative roles. A good prompt specifies the role to be filled. Contrastive prompting asks why one view handles a particular case better than its rival, rather than asking for a discussion of a topic in the abstract. Loveliness-sensitive prompting asks not for a conclusion but for a view that, if true, would explain more than its rival. Both target the explanatory virtues that make a philosophical answer worth reading. LLM philosophy improves when the prompt creates dialectical pressure: generic prompts invite generic continuations, and a philosophical prompt should create a space in which some argumentative move is needed, and then leave that move for the model to make. The creativity worry can be handled within the same continuum. Grammar correction is not interestingly creative. Generating a new objection or a new distinction may be. This fits a product-centred approach to creativity. Many accounts require novelty and value, and the present argument need not show that LLMs are creative agents in the fullest sense. It is enough that their outputs can contain novel and valuable philosophical structure. Nor does elicitation defeat creativity. Creative work often happens under constraints; a prompt can set a conceptual space without determining what is found within it. Elicited LLM philosophy can be creative at the level relevant to worth-readingness. The same point bears on the scarcity of good generic LLM texts. The corpus an LLM is trained on is the product of many rounds of philosophical criticism. A single completion does not reproduce that history. But a prompt can ask the model to stage, within the generated text, some of the operations that the corpus records across time: an objection pressed, a reply attempted, a distinction introduced under pressure. The human prompt supplies local conditions of elicitation; it need not supply the philosophical move. When the model generates that move, the output is LLM-produced in the sense relevant here. The absence of many great generic LLM texts is not the verdict it can appear to be. It reflects, at least in part, immature elicitation practices and the prevalence of generic prompting. If philosophy worth reading requires live alternatives and dialectical pressure, vague prompts will rarely elicit it. Elicitation does not defeat the claim that LLMs can produce philosophy worth reading: it shows that such production comes in degrees and depends on task conditions. The question to keep in view is whether the model has generated the philosophical structure in virtue of which the output is worth reading. The answer can be yes. LLMs do not produce worthwhile philosophy merely by being asked for “some philosophy”, and not every LLM-assisted text counts as LLM-produced in the relevant sense. But current models can generate philosophical moves that make a text worth reading, and where they do, the philosophical value is present in the product, and the product was produced by the LLM in the sense that matters here. ## Conclusion The four challenges treat what an LLM lacks as decisive for what its outputs can be. None of the absences they begin from is trivial. I have not denied those absences; I have argued that they do not, either singly or together, fix the philosophical status of the generated text. None of this settles whether what appears on the page is philosophy worth reading. A text repays philosophical attention when it changes what can be assessed in a dialectical context. Where an LLM output does that, the output is philosophy worth reading. What comes next is the same question we ask of any philosophical text: does the argument hold? ## Notes 1. The duplicate case is artificial in practice, but the artificiality is doing controlled work. By holding the textual product fixed, the case isolates the question whether causal history alters inferential structure. It does not. 2. The claim is not that every model trained on any philosophy-adjacent corpus will produce outputs with abductive structure. It is that the process described here is a plausible route by which current frontier models can do so. 3. The term “manipulative abduction” originates with Magnani, who introduced it for cases of hypothesis generation through the active construction of mental models. Zahavy uses it in this sense and applies it to Einstein’s elevator argument as a paradigm case in physics. 4. Iteration complicates attribution, but it does not automatically transfer production to the human. What matters is still whether the human supplies the philosophical move or elicits it. 5. Compare “Discuss whether physicalism is true” with a prompt asking why one version of physicalism can answer a specific objection that another cannot. The first asks for a survey. The second specifies a dialectical role. What matters is not the wording but the philosophical task imposed. [^1]: Consider the following example from the British science fiction novel: *The Hitchhiker’s Guide to the Galaxy*. At one point, humanity builds a computer and asks it for "the answer to the ultimate question of life, the universe, and everything"; after seven and a half million years, the computer returns ‘42’ — correct, apparently, but meaningless, given that mankind neglected to ask what the ultimate question is, they have no context in which to understand the answer. We will return to the question of questioning AI in Section 4, but for now we assume that worthwhile philosophy takes place at the level of the argument. [^2]: footnote saying that he calls it manipulative abduction. it should probably also explain why we might think go this as abduction as well as what we talked about in the previous section [^3]: A nice example in the footnote will be the feeling of understanding that is sometimes used as a way of motivating cognitive phenomenology. [^4]: mention that we assume this is some sort of mental imagery. a conscious experience in itself and derived, at least partly, from perceptual experience