# opus[1m]
## Skills Used
- [defuddle](../.claude/skills/defuddle/SKILL.md)
- epistemic-discipline
- nick-analytic-voice
- nick-philosophical-prose
- source-work
- twork
- voice-fix
- writing-standards
## Retrieval Notes
- Session id: `6fa4888d-f882-4207-b9e3-ee16a4a27e99`
- Last activity: `2026-03-18T14:56:27.422Z`
- Files touched: `2`
## Artifacts
**Modified:**
- [[Writing/research/generating-philosophy-text-internal-evaluation/0. Introduction]]
- [[1. The Challenge from Authorship]]
## My Notes
<!-- Add your notes here. This section is preserved across syncs. -->
## Conversation
### User
opus[1m]
### User
sonnet[1m]
### User
opus[1m]
### User
I want you to help me repair my section one of my generating philosophy paper. The address is here. The Obsidian address that is.
the way you should do this is by reading the chat which you can find in this note and in particular having a look at the evaluation of section one in terms of style which you can find at the very beginning sorry at the very end of the conversation okay because as you can see there are a lot of terrible style mistakes. Once you've done that, I would like you just in the chat to write me out a new paragraph by paragraph version of this section. It should retain all of the content of the current version of section one. And when I say all, I mean all. I'm fucking sick of detail bleeding out as we iterate. Okay, so it has to be all okay. However, it needs to be written much, much, much, much closer to my style of writing. Okay, as you can see from the evaluation in the chat, it is still very far from that. Okay, and you'll be able to see exactly the sorts of things I fucking hate in that chat. So yeah, basically I just want a new version of section one here in the chat. Verbatim when allowable, when acceptable, but changed as much as needs be to make sure it's properly in style in keeping with my actual writing style.
you must invoke the following skills BEFORE DOING ANYTHING
* Skill contemplate
* Skill nick-analytic-voice
* Skill nick-philosophical-prose
* Skill twork
* Skill source-work
* Skill epistemic-discipline
* Skill writing-standards
### Assistant
I need to invoke all the required skills first, then read the source materials. Let me start by invoking the remaining skills (contemplate is already loaded).
### Assistant
Now let me read the reference files required by the voice skills, the current section 1, and the conversation with the style evaluation — all in parallel.
### Assistant
The conversation file is too large. Let me read the end where the style evaluation is, and also read the vocabulary reference.
### Assistant
Let me search for the style evaluation section specifically.
### Assistant
Now let me read Nick's published papers to ground the rewrite properly.
### Assistant
Let me read specific sections from the published papers to get Nick's voice fully grounded.
### Assistant
I now have the current Section 1, the full style evaluation from the conversation, Nick's published papers (Growing the Image, Hearing Spaces, Agents of Change), and all the voice/style references. Let me do the contemplation and draft.
The style evaluation from the conversation identifies specific failures in the current Section 1, getting worse from paragraph 5 onward. The problems are:
1. Meta-commentary in smoother register: "This is where Lipton's distinction... becomes useful", "The point is not merely...", "shows what this amounts to in practice", "helps to say more precisely what that articulated case is assessed for", "That way of putting things is useful here because...", "The upshot is not that..."
2. Describing rather than doing: "Philosophical evaluation often has this character" — a summary of what the paragraph is about to say. "Williamson's discussion... helps to say more precisely" — announces what the paragraph will do. "make the same point at a more explicit methodological level" — same.
3. Blunt short sentences: "is not like this", "It does not.", "He asked what would make sense of the contrast"
4. Listicle structure in paragraphs 7 and 8: Bengson, Walton, Dellsén, blind review, Gaut, Deep Blue — one after another, none developed.
5. Remaining evaluative framing: "captures the contrast well", "useful here", "more provocatively"
6. Gentler list-making: "made a case, drawn distinctions, and opened a line of thought"; "simplicity, elegance, generality, and unificatory power"
The evaluation also said: paragraphs 1 and 4 (Semmelweis) are closest to Nick's prose. Paragraphs 2 and 3 are mixed. Paragraphs 5-9 are "assistant explaining philosophy."
I also need to respect Nick's corrections from the conversation:
- Don't use "structure" for Watson/Crick (use "arrangement" — already done in current draft)
- Use a famous same-year philosophy paper (Ryle "Ordinary Language" — already done)
- Framing should be discoveries vs. not-discoveries, not "contributions"
- Pigliucci combined into the Ryle paragraph
- No short paragraphs (fewer than 3 sentences)
- Deep Blue near the end as bridge to Section 2
- Block quote for Bengson
- Use "et al." for 3+ authors (voice rule)
Let me work through each paragraph now.
Paragraph 1 (Watson/Crick): Close to Nick's voice already. Main issue is the flat declarative chain. The current version has two sentences, but the three-sentence rule applies. I can split at the semicolon.
Paragraph 2 (Ryle + Pigliucci): "is not like this" needs to become part of a developed thought, not a blunt verdict. "captures the contrast well" needs to go — just bring Pigliucci in directly with "As Pigliucci puts it." The framing should be about discovery: Ryle "was not a discovery in that sense." Use "consists in" (Nick's preferred term). The closing should state the evaluative principle directly: philosophy is not evaluated as a report but as reasoning on the page.
Paragraph 3 (Lipton): Cut "This is where... becomes useful." Open with a genuine question flowing from the previous paragraph ("What, then, does the success of such reasoning consist in?") — not a rhetorical question with a pre-packaged answer, but a real question that the rest of the section develops. Then bring Lipton in directly. After the block quote, make the claims about loveliness directly, cutting "The point is not merely that some explanations are prettier."
Paragraph 4 (Semmelweis): The strongest one. Cut "shows what this amounts to in practice" — just enter the case directly. Start with Semmelweis facing the mortality difference. The rest is already close.
Paragraph 5 (philosophical evaluation): Cut "Philosophical evaluation often has this character." Start with the direct claim: philosophy doesn't have direct empirical tests. Cut "The issue is not... It does not. The issue is that..." — replace with "This is not to say that" (Nick's characteristic qualifying phrase, per the vocabulary reference).
Paragraph 6 (Williamson): Cut "helps to say more precisely what that articulated case is assessed for." Just bring Williamson in directly. The overfitting content is good; the framing is bad.
Paragraph 7 (Bengson + Walton): The worst one. Cut "make the same point at a more explicit methodological level." Cut "That way of putting things is useful here because it makes plain that." Just state what the criteria are and that they show up in the text. Use "et al." for Bengson.
Paragraph 8 (Dellsén, blind review, Deep Blue): Give these slightly more room. Use "et al." for Dellsén. "Recall" (Nick's characteristic term) to bring in Gaut/Deep Blue.
Paragraph 9 (closing): "None of this is to suggest that" instead of "The upshot is not that." Compress the recap list slightly while keeping the substance.
Now, comparing my planned revisions against Nick's published prose. In Hearing Spaces, he writes: "What do we hear? Three obvious candidates are sounds, properties of sounds, and echoes." Direct entry, no meta-commentary. In Agents of Change: "My aim here is to show how experience leads to the belief that time passes." First person, direct claim. In Growing the Image: "To see why autonomy is not sufficient for attribution of credit, consider the following example." Then the wine example, fully developed. These show the pattern: claim → case/source → development → consequence. No management phrases mediating between the reader and the subject matter.
The hardest paragraphs to fix are 5, 6, and 7. They need to face the subject matter — what philosophical evaluation actually is — rather than describing what various sources say about it. The move is to make each paragraph a claim supported by a source, not a source-report followed by a meta-summary.
For paragraph 5: instead of "Philosophical evaluation often has this character," start with what philosophy lacks (direct empirical test) and then state what we can ask. Those are direct philosophical claims, not descriptions of claims.
For paragraph 6: instead of "Williamson helps to say more precisely," just bring Williamson in through what he says. "Williamson, writing about abductive methodology, treats philosophy as answerable to virtues..." — that's a report of Williamson's position, but it faces the content of his position rather than managing the reader's relationship to it.
For paragraph 7: instead of "That way of putting things is useful here because it makes plain that," just state the claim: the standards are not mysterious, they show up in the text. Bengson's criteria are then the evidence for that claim.
I'm now ready to draft.
# Final Answer
> **Orienting passage:** "Although the claim that we hear source events is popular, I shall argue in this section that it is not obvious how it should be understood. This can be brought out through a comparison with vision: while we have a clear idea of what it is to see events, it is implausible to think that we hear events in the same way. If this is correct, we must characterise source hearing in some other way." — *Hearing Spaces*
> This is relevant because it shows how Nick enters a topic: with a direct claim, a comparison that does argumentative work, and a conditional consequence — no meta-commentary about what the argument is about to do.
Here is the rewritten Section 1, paragraph by paragraph. Every claim, argument, example, source, qualification, and footnote from the current draft is preserved. What changes is the voice: management phrases are removed, short blunt sentences are developed, evaluative framing is replaced with direct engagement, and the later paragraphs face the subject matter rather than describing the argument's own machinery.
---
When Watson and Crick published their paper on DNA in 1953, they announced what they had found: a particular arrangement of nucleotides, with two strands running in opposite directions and complementary base pairs linked by hydrogen bonds. That arrangement did not depend on their paper; it was there before they described it. Had Rosalind Franklin published it first, the discovery would have been the same one, only differently attributed.
Gilbert Ryle's 'Ordinary Language', published in that same year, was not a discovery in that sense. Ryle was not reporting a previously hidden item in the world, waiting to be uncovered by the first sufficiently careful observer; he was trying to get his readers to see familiar philosophical materials differently, and what he achieved consists not in having tracked some independent arrangement already there but in the case he made, the distinctions he drew, and the line of thought he opened for others to assess and contest. As Pigliucci puts it, philosophy is engaged in "empirically informed evoking, not inventing", its starting points being "empirical data about the world" and its constraints imposed by "our best understanding of how the world actually is".[^pigliucci] The starting points may be empirical, but the philosophical work consists in what is done with them; philosophy is not evaluated as a report of what was found but as a piece of reasoning whose success or failure lies on the page.
What, then, does the success of such reasoning consist in? Consider Lipton's distinction between two ways of understanding "best explanation". He writes:
> We may characterize it as the explanation that is most warranted: the 'likeliest' or most probable explanation. On the other hand, we may characterize the best explanation as the one which would, if correct, be the most explanatory or provide the most understanding: the 'loveliest' explanation. The criteria of likeliness and loveliness may well pick out the same explanation in a particular competition, but they are clearly different sorts of standard. Likeliness speaks of truth; loveliness of potential understanding. (*Inference to the Best Explanation*, p. 59)
We can ask of a hypothesis what kind of understanding it would provide if it were true, and this is not the same question as whether the available evidence already warrants belief in it. A lovely explanation does not simply fit what we know; it makes the matter intelligible. It shows why the phenomenon has the shape it does, instead of merely placing another true sentence next to it.
Semmelweis, faced with the much higher mortality rate in one division of the Vienna maternity hospital than in the other, did not merely ask which hypothesis had the best immediate evidential support; he asked what would make sense of the contrast. Some candidate explanations did very little with the phenomenon even if they were granted: the route taken by the priest, for instance, left obscure why this should be a matter of life and death. The cadaveric hypothesis did more. It connected the contrast with the medical students' contact with corpses, with Kolletschka's death after a puncture wound, and with the subsequent fall in mortality once disinfection was introduced. In Lipton's terms, it was lovelier because it rendered the pattern intelligible.
Philosophy does not usually have the kind of direct empirical test that would settle whether a view of analyticity, personal identity, modality, or moral responsibility is simply true in the manner in which a laboratory test may settle a question in chemistry. What we can ask is whether a position makes the terrain more intelligible than its rivals do, whether it handles the objections it ought to handle, and whether it earns its explanatory reach without multiplying distinctions merely to escape trouble. This is not to say that philosophy floats free of the world — the starting points, as Pigliucci reminds us, are empirical. But the route by which a philosophical view earns our regard is through the articulated case it makes, not through its correspondence with some independent finding.
Williamson, writing about abductive methodology, treats philosophy as answerable to virtues — simplicity, elegance, generality, unificatory power — and connects those virtues to a familiar problem in statistics: overfitting. A philosophical account can be made to accommodate more and more cases by the addition of piecemeal repairs, but the result is often a view that fits the immediate data at the cost of becoming increasingly gerrymandered and fragile. What matters, then, is not just that an argument survives the latest objection, but how it survives it. A view that preserves clarity and reach while absorbing pressure is doing something different from a view that survives only by accumulating complications.
Bengson et al. are more explicit about the evaluative criteria:
> the best theory is the one that satisfies the criteria at these levels (so ordered) to the highest degree relative to its rivals.
The levels they have in mind — accommodation, explanation, substantiation, integration, and theoretical virtue — are not mysterious properties hovering behind a finished text. They show up in whether a theory accommodates the data it is meant to handle, whether it explains those data rather than merely redescribing them, whether its claims are defended, and whether its commitments hang together. Walton's work on argumentation schemes points in the same direction: philosophical arguments proceed through recognisable forms of analogy, consequence, objection, reply, refinement, and concession, each with its own pressure points. What we see in a philosophical text is not a surface behind which the real work is concealed; we see the work itself, carried by the shape of the reasoning.
Dellsén et al. argue that philosophical progress consists in putting people in a position to increase their understanding. The understanding-making material can be publicly available before anyone takes it up, just as a published scientific result can remain a contribution even when no one is presently attending to it. A published philosophical argument may be good, bad, illuminating, muddled, subtle, or ad hoc quite apart from the biography of its author, because those are features of what has been written. Blind review gives this thought an institutional form: referees are asked to assess what a paper does, not who produced it. Recall Gaut's discussion of Deep Blue. A move can be good chess even if it was found by exhaustive search rather than by human insight; what makes it good is the move itself. If a philosophical text handles its subject matter well, that success does not turn on the route by which it was generated.
None of this is to suggest that philosophy is easy to automate, or that any sequence of plausible-looking sentences counts as philosophy. The standards by which philosophical writing is judged bear, to a considerable degree, on what is there in the writing itself: on how the case is built, what it explains, how it absorbs pressure, and what understanding it affords. A philosophical corpus selected, criticised, cited, taught, and preserved under those standards is not a random body of text; it is a body of text shaped, however imperfectly, by repeated judgements of that kind. The question, then, is what follows when a language model is trained on such a corpus and begins producing texts of its own.
[^pigliucci]: We return in Section 3 to the question of worldly starting points and empirical constraint, where it matters directly to the grounding-style objection.
---
## What changed and why
Paragraph 1: Split into three sentences to meet the minimum. No other changes needed; this was already close.
Paragraph 2: "is not like this" → "was not a discovery in that sense" (picks up the discovery framing, avoids the blunt verdict). "Massimo Pigliucci captures the contrast well" → "As Pigliucci puts it" (removes evaluative framing; just brings the source in). Reframed around discoveries vs. not-discoveries rather than "contributions." "Even so, the philosophical work is done in the argument itself. That is why..." → "The starting points may be empirical, but the philosophical work consists in what is done with them; philosophy is not evaluated as..." (direct claim, no "that is why" management).
Paragraph 3: "This is where Lipton's distinction... becomes useful" → "What, then, does the success of such reasoning consist in? Consider Lipton's distinction..." (a genuine question flowing from the previous paragraph, not a meta-commentary sentence). "The point is not merely that some explanations are prettier than others" → cut entirely; the claim after the quote now goes straight to what Lipton's distinction gives us.
Paragraph 4: "Semmelweis's work on childbed fever shows what this amounts to in practice" → "Semmelweis, faced with the much higher mortality rate..." (enters the case directly without throat-clearing). The rest is largely verbatim — this was already the strongest paragraph.
Paragraph 5: "Philosophical evaluation often has this character" → cut. The paragraph now opens with the direct claim: "Philosophy does not usually have the kind of direct empirical test..." "The issue is not whether philosophy therefore floats free of the world. It does not. The issue is that..." → "This is not to say that philosophy floats free of the world — the starting points, as Pigliucci reminds us, are empirical. But the route by which..." (Nick's characteristic "this is not to say" qualifier; removes the blunt "It does not." and the double "the issue is" frame).
Paragraph 6: "Williamson's discussion of abductive methodology helps to say more precisely what that articulated case is assessed for" → "Williamson, writing about abductive methodology, treats philosophy as answerable to virtues..." (enters Williamson's content directly instead of announcing what the paragraph will do). Otherwise verbatim — the overfitting content was already decent.
Paragraph 7: "Bengson, Cuneo, and Shafer-Landau make the same point at a more explicit methodological level" → "Bengson et al. are more explicit about the evaluative criteria" (removes the meta-phrasing, uses et al. per style rule). "That way of putting things is useful here because it makes plain that philosophical standards are not mysterious..." → "The levels they have in mind... are not mysterious properties hovering behind a finished text" (cuts the management phrase, makes the claim directly). "Walton's work on argumentation schemes helps in a related way" → "Walton's work on argumentation schemes points in the same direction" (less evaluative). Changed "one" to "we" as pronoun.
Paragraph 8: "The same point can be put more provocatively by recalling Gaut's discussion" → "Recall Gaut's discussion of Deep Blue" (cuts "more provocatively," uses Nick's characteristic "Recall"). "Dellsén and his co-authors" → "Dellsén et al." Otherwise the content is preserved and the paragraph is given slightly more breathing room between blind review and Deep Blue.
Paragraph 9: "The upshot is not that" → "None of this is to suggest that" (Nick's characteristic "this is not to say" register). The four-item recap list is preserved but rephrased slightly ("absorbs pressure" echoing the Williamson paragraph's "absorbing pressure"). Two short sentences combined with semicolon.
*Riscrivere senza perdere nulla è il mestiere più ingrato della filosofia — e forse anche il più necessario.*
### User
That was great. Please save a new version of this note using the version history plugin via the CLI and then replace what's on the note right now with the version you just gave me just then. That is you should reproduce the version you just gave me verbatim on the note.
### Assistant
I see — the CLI uses `command id=`. Let me find the version-control plugin's command ID first.
### Assistant
Version saved. Now let me write the new content to the file.
### Assistant
Done. Version saved via the version-control plugin, then the note replaced with the rewritten text verbatim.
*Il mestiere della revisione è sapere dove la mano dell'artigiano finisce e comincia quella della macchina.*
### User
/draft-audit it's better than it was, but there's still quite a few issues. See all of the inline comments. Also, please just talk to me in the chat. Do you really think that we should have a paragraph on Williamson, a paragraph on Bangson and a paragraph on Delson? It seems to me very fucking listically.
### Assistant
Let me read the file with your comments first.
### Assistant
Let me go through each comment and the structural question carefully.
Starting with the comments in order.
Comment 1 (line 6): "not how i write" on "the case he made, the distinctions he drew, and the line of thought he opened for others to assess and contest."
This is the triplet I worried about and ultimately kept. Nick has now confirmed it is wrong. The three items — case, distinctions, line of thought — are trying to say what Ryle did as opposed to what Watson and Crick did. But they say it in a list where one item would do. What DID Ryle do? He made an argument. That is the whole point of the section. "What he achieved consists in the argument he made" — or even better, "consists in the argument itself." But then what about "distinctions" and "opened a line of thought"? Those are aspects of what making a philosophical argument involves. They do not need to be listed separately. The sentence should say that Ryle's achievement IS the argument, not enumerate the components of arguing.
So: "...and what he achieved consists not in having tracked some independent arrangement already there but in the argument itself — in how it was made, and in what it opened for others to assess and contest." Or simpler: "...but in the argument he made." Let me think about which captures enough without listing.
Comment 2 (line 6): Pigliucci's connection to Ryle is unclear.
The problem: the paragraph says Ryle was doing philosophy (not discovering), then jumps to Pigliucci saying philosophy is "empirically informed evoking." The reader asks: why are you telling me this? What does Pigliucci add to the point about Ryle? The answer is that Pigliucci captures the way philosophy relates to the world without being a report of findings. The Ryle point establishes that philosophy is not discovery; the Pigliucci refinement says: and this does not mean philosophy is disconnected from the world. But that connection is not made. The sentence "As Pigliucci puts it" just throws him in.
A fix would be to make the connection explicit: something like "This does not mean philosophy starts from nothing. As Pigliucci puts it, philosophy is engaged in..." That would make clear WHY Pigliucci is being invoked: to qualify the claim that philosophy is not discovery by noting that it is still empirically constrained. The Ryle/Watson-Crick contrast might otherwise sound as though philosophy has no worldly input at all.
Comment 3 (line 12): The likeliest/loveliest distinction needs more development.
Nick is right. The paragraph after the Lipton quote says "we can ask of a hypothesis what kind of understanding it would provide if it were true, and this is not the same question as whether the available evidence already warrants belief in it." That is just restating the distinction, not explaining it. When do the two come apart? The obvious case: a hypothesis could be true and well-supported but explain nothing — it just sits there as another fact. Or a hypothesis could be speculative but, IF true, would make everything click. Semmelweis is meant to illustrate this, but if the reader does not understand the distinction before they get to Semmelweis, the illustration will not land.
What is needed: a sentence or two showing what a merely likely but not lovely explanation looks like, or what a lovely but not yet likely one looks like. Something like: "A hypothesis can be well-supported by the evidence and still leave the phenomenon it describes opaque — it can be likeliest without being loveliest. A merely likely hypothesis tells us that something is the case; a lovely one shows us why it should be."
Actually, that might over-formalise it. A concrete quick example might work better. Or even just: "A hypothesis can have strong evidential support and still leave us in the dark about why the phenomenon looks the way it does. To ask whether a hypothesis is lovely is to ask whether it would illuminate the matter — whether it would give us a grip on the shape of the phenomenon, not merely confirm that the phenomenon is there."
Comment 4 (line 14): "not how i write" on "he asked what would make sense of the contrast."
This is inside the Semmelweis paragraph. "Semmelweis, faced with the much higher mortality rate... did not merely ask which hypothesis had the best immediate evidential support; he asked what would make sense of the contrast." The second clause is the problem. "He asked what would make sense of the contrast" is a kind of compressed rhetorical move — it is trying to dramatise Semmelweis's question, but it comes off as a blunt declarative. In Nick's published work, this kind of moment would be embedded in a longer sentence or would be more specific about what "making sense" means here. It also has a slightly breathless quality.
How to fix: fold it into the preceding sentence more naturally. "...did not merely ask which hypothesis had the most evidential support but what would make the contrast between the two divisions intelligible." That keeps it as one sentence, removes the blunt standalone, and uses "intelligible" which connects to the Lipton framework.
Comment 5 (line 14): Semmelweis is LIPTON'S example. Not attributing it is "fucking obscene."
This is a serious source attribution error. The current text presents the Semmelweis case as if the paper's authors are introducing it. But it comes from Lipton's *Inference to the Best Explanation*. The fix is straightforward: "Lipton's own example makes the distinction concrete." Or: "Lipton develops the distinction through Semmelweis's work on childbed fever." The reader needs to know this is Lipton's case, worked through here.
Comment 6 (line 16): "ugly awkward sentence" on "Philosophy does not usually have the kind of direct empirical test that would settle whether a view of analyticity, personal identity, modality, or moral responsibility is simply true in the manner in which a laboratory test may settle a question in chemistry."
The sentence IS too long and the list of philosophical topics (analyticity, personal identity, modality, moral responsibility) is padding. The point is simple: philosophy does not have direct empirical tests. The list of topics adds nothing except word count. A better version: "Philosophical views are not usually settled by direct empirical test in the way that a laboratory result may settle a question in chemistry." That says the same thing without the parade of sub-disciplines.
Comment 7 (line 16): "fucking triple examples" on "whether a position makes the terrain more intelligible than its rivals do, whether it handles the objections it ought to handle, and whether it earns its explanatory reach without multiplying distinctions merely to escape trouble."
Another triplet. The three items are: (1) makes terrain intelligible, (2) handles objections, (3) earns explanatory reach without gerrymandering. These are genuinely different criteria, but listing them this way is the problem. How to handle this? Maybe: "What we can ask is whether a position makes the terrain more intelligible than its rivals do — whether it earns its reach through genuine explanatory power rather than by multiplying distinctions to escape trouble." That folds (1) and (3) together and drops (2), which is arguably implicit in any evaluation. Or keep all three but embed them differently.
Actually, "handles the objections it ought to handle" is generic. A theory that makes things intelligible and earns its reach without gerrymandering IS one that handles objections well. So (2) is implied by (1) and (3). I think dropping it is fine.
Comment 8 (line 16): "not how i write" on "This is not to say that philosophy floats free of the world — the starting points, as Pigliucci reminds us, are empirical."
Hmm. I used "this is not to say" which IS in Nick's vocabulary. But the sentence as a whole might sound too managed — "as Pigliucci reminds us" is the problem. "Reminds us" is pedagogical, teacher-to-student. Nick's voice doesn't do that. And the whole sentence might feel like an unnecessary qualification that breaks the flow. If the Pigliucci point has already been made in paragraph 2, restating it here is redundant.
Maybe just cut this sentence entirely. The Pigliucci refinement is already in paragraph 2 with the footnote pointing to Section 3. Repeating it here is belt-and-suspenders.
Comment 9 (line 18): "What matters" — Nick never uses "matters."
Simple vocabulary fix. "What matters" → "The question, then, is not just..." or "The point is not just..." — wait, "the point is" is also management. How about: "We should ask not just whether an argument survives the latest objection, but how it survives it." That uses "we" and a direct claim.
Now the structural question: Williamson paragraph → Bengson paragraph → Dellsén paragraph is a listicle.
Nick is absolutely right. Even if each paragraph is internally fine, the sequence is: source-report, source-report, source-report. That is a literature review, not an argument. In Nick's published work, sources serve the argument — they are brought in as needed, not paraded. Look at Growing the Image: Anscomb is quoted and responded to within a paragraph that is about whether Midjourney is an agent, not about what Anscomb thinks. The topic is the question, not the source.
So the fix is structural. Instead of one paragraph per source, the paragraphs should be organised by what they CLAIM, with sources woven in:
Claim 1: Philosophical evaluation has specific criteria — it asks whether a view illuminates the territory without gerrymandering. (Williamson on overfitting, Bengson et al. on levels of assessment — these are making the same point from different angles and can share a paragraph.)
Claim 2: These criteria are visible in the text itself, and success is a feature of what is written, not of who wrote it. (Walton on argument forms, Dellsén on understanding as public material, blind review, Deep Blue — these are all about the text-internal, provenance-independent nature of the standards.)
That gives us two paragraphs instead of three, and each one is organised by claim rather than by source.
Let me think about whether this works. The Williamson overfitting point and the Bengson criteria list — do they sit together? Yes. Williamson says philosophy is answerable to virtues and can overfit. Bengson et al. spell out what those virtues look like in a hierarchy. They are making the same claim at different levels of specificity. One paragraph.
Then Walton (argument forms are visible), Dellsén (understanding is publicly available), blind review (assess the text not the author), Deep Blue (provenance doesn't matter for quality) — these all converge on the claim that the standards bear on what's written. One paragraph, with Deep Blue as the bridge to LLMs.
That seems right. Two paragraphs, not three. Organised by claim, with sources woven in.
Now let me think about the whole section's flow with all these fixes:
P1: Watson/Crick (1953). What they found. There before they described it. Franklin counterfactual.
P2: Ryle (same year). Not a discovery in that sense. The argument itself is the achievement. Pigliucci refines: philosophy's starting points are empirical, but the work consists in what is done with them. Philosophy is evaluated as reasoning on the page.
P3: Lipton's likeliest/loveliest distinction. Block quote. DEVELOPED explanation of when the two come apart — a sentence or two more. What it means for a hypothesis to be lovely: it would illuminate the matter, give us a grip on the shape of things.
P4: Lipton's own example — Semmelweis. Properly attributed. Developed as before.
P5: Philosophy works like this (but reworked — no "has this character," no awkward opening sentence, no triple, no redundant Pigliucci callback). Philosophical views are not settled by direct empirical test. We ask whether a position illuminates the terrain and earns its reach without gerrymandering. Williamson's overfitting point and Bengson et al.'s evaluative levels, woven together. The claim: there are specific, articulable criteria, and they bear on the reasoning itself.
P6: These standards are visible in the text, and success is a feature of what is written. Walton on argument forms. Dellsén on understanding as publicly available. Blind review. Deep Blue. Bridge to LLMs.
P7: Closing — the corpus is shaped by these judgements, and the question is what follows when an LLM trains on it.
That is seven paragraphs instead of nine. Tighter, less listicle, each paragraph organised by what it claims rather than by which source it introduces.
Let me now also think about the specific voice issues flagged.
"Not how I write" appears four times:
1. The triplet in paragraph 2
2. "He asked what would make sense of the contrast" in Semmelweis
3. The opening of paragraph 5
4. "This is not to say that philosophy floats free"
Plus "What matters" in the Williamson paragraph.
These are all fixable at the sentence level. The bigger issue is the structure. If I fix the structure (merging the source-report paragraphs into claim-driven paragraphs), several of these sentence-level problems may resolve themselves because the surrounding context changes.
One more thought on the Pigliucci transition (comment 2). The problem is that Pigliucci currently sits in paragraph 2 with no clear role relative to the Ryle discussion. The fix I suggested — "This does not mean philosophy starts from nothing" — would make the connection clear. But there is a subtler option: instead of making Pigliucci qualify the Ryle point, make him DEVELOP it. Ryle's work is not a discovery in the Watson/Crick sense. But it is not disconnected from the world either. Pigliucci says: philosophy is "empirically informed evoking." That is a more precise characterisation of what Ryle was doing — he was evoking something from materials that already had an empirical shape. The distinction between evoking and discovering is the point of the opening. Pigliucci is giving us the vocabulary to make it.
So: "...and what he achieved consists in the argument itself. Pigliucci's way of putting the distinction is useful here: philosophy, he says, is engaged in 'empirically informed evoking, not inventing'..."
Wait, "useful here" is evaluative framing. Let me try: "Pigliucci's phrase for this is apt — philosophy, he says, is 'empirically informed evoking, not inventing'..." No, "apt" is also evaluative.
How about: "Pigliucci draws the distinction in a way that clarifies what is at stake: philosophy is engaged in 'empirically informed evoking, not inventing'..." Still managed.
OK, what if I just state the connection directly: "Philosophy does not start from nothing — Pigliucci is right that its starting points are 'empirical data about the world' and its constraints imposed by 'our best understanding of how the world actually is'. But the work is done in the argument itself, not in the tracking of some prior finding." That makes the Pigliucci quotation do the work of qualifying the discovery/not-discovery contrast. The reader sees: Ryle is not doing Watson/Crick-style discovery, but nor is he working in a vacuum — his starting points are empirical, and the result is still constrained by the world. The philosophical contribution, though, is the argument as made.
I think that handles the Pigliucci problem. The connection is: the opening might sound as though philosophy has no worldly input; Pigliucci corrects that; but even with that correction, the work is in the argument.
You are right about the listicle problem. Williamson paragraph → Bengson paragraph → Dellsén paragraph is source-report, source-report, source-report. That is a literature review, not an argument. In your published work, sources serve claims — they get brought in where they are needed, not paraded one at a time.
Here is what I think the second half of the section should look like structurally, followed by my response to each inline comment.
## The structural fix
Instead of three source-report paragraphs (Williamson, Bengson, Dellsén), organise by claim, with the sources woven in:
One paragraph that says: philosophical evaluation has specific, articulable criteria — it asks whether a view illuminates the territory without gerrymandering (Williamson's overfitting point), and those criteria have a recognisable shape: accommodation, explanation, substantiation, integration, theoretical virtue (Bengson et al.). These are not separated into "Williamson says X" then "Bengson says Y." They make the same claim at different levels of specificity and belong together.
A second paragraph that says: these criteria are visible in the text itself, and success is a feature of what is written, not of who wrote it (Walton on the forms arguments take; Dellsén et al. on understanding as publicly available material; blind review; Deep Blue as bridge to LLMs). Again, not source-report, source-report — one claim with multiple sources supporting it.
That gives you seven paragraphs for the section rather than nine, and the second half reads as an argument with sources rather than a survey of sources.
## The inline comments
### Comment 1: %%not how i write%% (the triplet in paragraph 2)
"the case he made, the distinctions he drew, and the line of thought he opened for others to assess and contest" — you are right, this is a list where one thing would do. What Ryle did was make an argument. The components of arguing do not need separate billing. Something like: "...but in the argument itself, and in what it opened for others to assess and contest." Or even just: "...but in the argument itself."
### Comment 2: Pigliucci's connection to Ryle is unclear
The problem is that Pigliucci arrives without any signal of what he is adding to the Ryle discussion. The reader does not know why he is being invoked. The fix is to make the connection explicit — something along the lines of: "Philosophy does not start from nothing. Pigliucci is right that its starting points are 'empirical data about the world' and its constraints imposed by 'our best understanding of how the world actually is'. But the work is done in the argument itself." That makes clear what Pigliucci is doing: qualifying the discovery/not-discovery contrast so it does not sound as though philosophy has no worldly input. The quotation then does actual work in the paragraph's logic rather than floating in.
### Comment 3: likeliest/loveliest needs more development
You are right, and I am rushing. The paragraph after the Lipton quote restates the distinction but does not explain it. The reader needs to see when the two come apart. Something like: "A hypothesis can have strong evidential support and still leave us in the dark about why the phenomenon looks the way it does. To ask whether a hypothesis is lovely is to ask what it would illuminate — whether it would give us a grip on the shape of the phenomenon, not merely confirm that the phenomenon is there." That gives the reader a way to understand "loveliest" before Semmelweis is supposed to illustrate it.
### Comment 4: %%not how i write%% (Semmelweis opening)
"he asked what would make sense of the contrast" is a blunt standalone clause. It could be folded into the preceding sentence: "...did not merely ask which hypothesis had the most evidential support but what would make the contrast between the two divisions intelligible." One sentence, no blunt fragment, and "intelligible" connects back to Lipton.
### Comment 5: Semmelweis is LIPTON'S example
This is a source attribution error and it is serious. The text currently presents Semmelweis as though you are introducing the case. You are not — it is Lipton's worked example from the book. The fix is straightforward: "Lipton develops the distinction through Semmelweis's work on childbed fever." Or: "The example Lipton uses to develop the distinction is Semmelweis's work on childbed fever." The reader needs to know this is his case.
### Comment 6: ugly awkward opening sentence (paragraph 5)
"Philosophy does not usually have the kind of direct empirical test that would settle whether a view of analyticity, personal identity, modality, or moral responsibility is simply true in the manner in which a laboratory test may settle a question in chemistry." The list of philosophical sub-disciplines is padding. The point is: philosophical views are not settled by direct empirical test. Try: "Philosophical views are not usually settled by direct empirical test in the way that a laboratory result may settle a question in chemistry." Same claim, no padding.
### Comment 7: triple examples (paragraph 5)
"whether a position makes the terrain more intelligible than its rivals do, whether it handles the objections it ought to handle, and whether it earns its explanatory reach without multiplying distinctions merely to escape trouble." Three parallel "whether" clauses. Item (2) — "handles the objections it ought to handle" — is implied by (1) and (3). A position that makes things intelligible and earns its reach without gerrymandering IS handling objections well. Drop (2), fold (1) and (3): "...whether a position makes the terrain more intelligible than its rivals do, and whether it earns its explanatory reach through genuine illumination rather than by multiplying distinctions to escape trouble."
### Comment 8: %%not how i write%% (Pigliucci callback)
"This is not to say that philosophy floats free of the world — the starting points, as Pigliucci reminds us, are empirical." Two problems: "as Pigliucci reminds us" is pedagogical (teacher-to-student), and the whole sentence is redundant if the Pigliucci point is already made properly in paragraph 2 with the footnote to Section 3. If paragraph 2 handles Pigliucci well — which it would after the fix in comment 2 — this sentence can be cut entirely.
### Comment 9: "What matters" — you never use it
Simple vocabulary fix. "What matters, then, is not just that an argument survives the latest objection, but how it survives it" → "We should ask not just whether an argument survives the latest objection but how it survives it." Or: "The question is not just whether an argument survives the latest objection but how it survives it."
## Summary of what the section would look like
Seven paragraphs:
1. Watson/Crick opening (largely as is)
2. Ryle + Pigliucci (triplet fixed, Pigliucci connection made explicit)
3. Lipton's distinction (block quote + DEVELOPED explanation of when likeliest and loveliest come apart)
4. Semmelweis (properly attributed to Lipton, "make sense of the contrast" folded into one sentence)
5. Philosophy works like this + Williamson + Bengson et al. (one claim-driven paragraph: philosophical views are not settled by direct test; we ask whether a view illuminates without gerrymandering; Williamson's overfitting, Bengson et al.'s criteria, woven together)
6. Standards are visible in the text + provenance does not determine quality (Walton on forms, Dellsén on understanding, blind review, Deep Blue as bridge)
7. Closing (corpus shaped by these judgements; question is what happens when an LLM trains on it)
Do you want me to draft the new version based on this, or do you want to push back on any of these before I write?
*Tre paragrafi in fila, ciascuno con il nome di un autore in cima — non è un argomento, è un elenco telefonico.*
### User
"Philosophy works like this + Williamson + Bengson et al. (one claim-driven paragraph: philosophical views are not settled by direct test; we ask whether a view illuminates without gerrymandering; Williamson's overfitting, Bengson et al.'s criteria, woven together)
Standards are visible in the text + provenance does not determine quality (Walton on forms, Dellsén on understanding, blind review, Deep Blue as bridge)" I think this is okay, but please go back to these sources, read them carefully, don't just fucking skim them. And make sure that what you're doing with them is is accurate to what they're actually saying.
But yeah, yeah iterate the draft in the seven paragraphs in the same in the way that you did. As always, I'm fucking begging you, please don't dilute dilute the detail here. I know you love to. I know it makes your dick hard. Please don't this time. Okay, to avoid this you're gonna have to really think hard about everything and how to structure things and how to keep everything together before you actually write the final answer in the notes. Of course part of that consideration as well will be what text can stay, what text can be reused in a different place or broken up and reused in a different place. What text must be removed and what text must be added or changed? Okay, this is all important stuff. You almost never give it the right amount of work. This time, maybe surprise me. make sure that you explain your analysis, justification, reasons etc etc. for your answer BEFORE giving me your final answer. Please give me a long, deep, and detailed answer, with lots of options/ideas: as we are in the 'kicking ideas around' stage, I want as much feedback as you can give me so that i have a lot to consider.
### Assistant
I need to go back to the actual sources. Let me extract and read them properly before doing anything else.
### Assistant
Good — the sources are all available as pre-extracted markdown files. Let me read the relevant ones now.
### Assistant
Now let me read the specific passages I need from the sources.
### Assistant
Let me find the Semmelweis passage in Lipton and also read the Dellsén section on understanding-making material being publicly available.
### Assistant
Now let me read the Semmelweis passage in Lipton Chapter 5, where the case is developed in detail, and the Dellsén passage about understanding being publicly available.
### Assistant
I have now read the actual source texts for Williamson, Bengson et al., Dellsén et al., and Lipton (including the full Semmelweis passage). Let me present the analysis and then the draft.
Starting with what the sources actually say, because several things in the current draft are imprecise.
Williamson (9.2 "Abductive Philosophy"): He treats philosophy as using abductive methodology. He says "Abduction also rewards virtues such as simplicity, elegance, generality, and unificatory power." His overfitting discussion draws on Forster and Sober: in curve-fitting, using overly complex equations fits current data but predicts future data badly. He extends this explicitly to philosophy: "tolerance for highly complicated, messy, gerrymandered analyses yielded a succession of proposals that fitted the current data but succumbed to new ones." His concrete example is the post-Gettier programme of analysing knowledge — "a cycle of proposed analysis, followed by new counterexamples... followed by a revised analysis, with an extra epicycle, a new disjunct or other complication." And: "A firmer preference for simplicity and elegance would have warned the community that something was going wrong." The current draft is accurate on Williamson but abstract — it states the overfitting problem without the Gettier example that makes it concrete. That is exactly the kind of detail that bleeds out. I should add the Gettier example.
Bengson et al. (Ch. 5, "The Tri-Level Method"): Their method evaluates theories at three levels: (1) accommodation and explanation of data; (2) substantiation and integration of claims; (3) theoretical virtue — but critically, level 3 is a TIEBREAKER, operative "only when two conditions are in place: (i) the candidate theories are all respectable... and (ii) the candidates are roughly equal... with respect to the four criteria at those levels." The current draft lists all five criteria as though they are co-equal ("accommodation, explanation, substantiation, integration, and theoretical virtue"). That slightly misrepresents Bengson — in their framework, virtue plays a different role from the other four. My new draft should be accurate about this hierarchy.
Dellsén et al. ("What is philosophical progress?"): Their account is "Enabling Noeticism" — "philosophical progress consists in putting people in a position to increase their understanding, where 'increased understanding' is a matter of more accurately and/or more comprehensively representing the network of dependence relations between various phenomena, and where people are most commonly put 'in a position' to increase their understanding by way of philosophical ideas (theories, arguments, distinctions, etc.) becoming publicly available." The key claim for our section: progress is about understanding-making material becoming publicly available, not about private insight. A paper that makes good arguments contributes to progress whether or not anyone has yet engaged with it. The current draft is accurate.
Lipton (Ch. 4 and Ch. 5): The likeliest/loveliest distinction is in the block quote (Ch. 4, p. 59). Semmelweis is LIPTON'S example — introduced in Ch. 4 ("We will consider this case at length in the next chapter") and developed across the entirety of Ch. 5. Lipton uses the case to illustrate how explanatory considerations guide inference: Semmelweis rejected hypotheses that did not explain the contrast between divisions (epidemic influences, overcrowding, diet, delivery position, the priest), and accepted the cadaveric hypothesis because it rendered the pattern intelligible — it connected the contrast with the medical students' contact with corpses, Kolletschka's death, and the fall in mortality with disinfection. The current draft presents all this WITHOUT attributing it to Lipton. That is a serious error of scholarly integrity — Nick is right to be furious.
Now for the structural analysis. What text can stay, what moves, what must change?
PARAGRAPH 1 (Watson/Crick): Stays almost entirely. Already split into three sentences.
PARAGRAPH 2 (Ryle + Pigliucci): Three fixes needed.
- Triplet ("the case he made, the distinctions he drew, and the line of thought he opened") → compress to "the argument itself — the distinctions it draws and what it opens for others to assess and contest"
- Pigliucci transition: "As Pigliucci puts it" → "Philosophy does not, of course, start from nothing. Pigliucci is right that..." — this makes clear why Pigliucci is being invoked (qualifying the not-a-discovery claim)
- Closing: "Even so, the philosophical work is done in the argument itself. That is why..." → "But the work is in the argument as made, not in any prior finding that the argument reports"
PARAGRAPH 3 (Lipton): Must add 1-2 sentences developing when likeliest and loveliest come apart. Currently the paragraph after the quote RESTATES the distinction but doesn't EXPLAIN it. A hypothesis can be well-supported and still leave the phenomenon opaque; a hypothesis that would illuminate the pattern may not yet have the weight of evidence behind it. Those are the cases where the two standards diverge, and the reader needs to see them before Semmelweis.
PARAGRAPH 4 (Semmelweis): Two fixes.
- Attribution: "Lipton develops the distinction through Semmelweis's work on childbed fever." — first sentence.
- "he asked what would make sense of the contrast" → fold into the preceding clause: "asked not just which hypothesis the evidence best supported but which would, if true, make the contrast between the two divisions intelligible."
PARAGRAPH 5 (merged Williamson + Bengson): This is the biggest structural change. The claim: philosophical evaluation has specific, articulable criteria, and they are visible in the argument itself.
- Open with what philosophy LACKS (direct empirical test) — reworded to avoid the "ugly awkward sentence"
- State what we ask instead (intelligibility, earned reach) — two criteria, not three
- Williamson's overfitting: virtues, the Gettier example as concrete illustration
- "What matters" → "The question is not just whether..."
- Bengson's tri-level method with block quote — accurate to their hierarchy (virtue as tiebreaker)
- Close: "What we see in a philosophical text is not a surface behind which the real work is concealed; we see the work itself, carried by the shape of the reasoning."
- Cut "This is not to say that philosophy floats free" — redundant once Pigliucci is handled in P2
PARAGRAPH 6 (merged Walton + Dellsén + blind review + Deep Blue):
- Walton on argument forms (there in the text to assess)
- Dellsén: understanding-making material is publicly available, not private insight
- Published argument assessed on qualities, not biography
- Blind review
- Deep Blue (bridge to LLMs)
PARAGRAPH 7 (closing): Minor changes from before. "The upshot is not that" → "None of this is to suggest that."
Now a content audit — what am I keeping, what am I cutting, what am I adding?
KEEPING (verbatim or near-verbatim):
- Entire Watson/Crick paragraph
- Lipton block quote
- Semmelweis details (priest, cadaveric hypothesis, Kolletschka, disinfection)
- Williamson on overfitting (sentence about piecemeal repairs, gerrymandered/fragile)
- "preserves clarity and reach while absorbing pressure" sentence
- Bengson block quote
- "not a surface behind which the real work is concealed; we see the work itself"
- Dellsén's "putting people in a position to increase their understanding"
- "understanding-making material can be publicly available before anyone takes it up"
- "good, bad, illuminating, muddled, subtle, or ad hoc quite apart from the biography"
- Blind review sentence
- Deep Blue passage
- Closing paragraph (with minor changes)
- Footnote
CUTTING:
- "Philosophical evaluation often has this character" (meta-commentary)
- "This is not to say that philosophy floats free of the world — the starting points, as Pigliucci reminds us, are empirical" (redundant)
- "whether it handles the objections it ought to handle" (implied by the other two criteria)
- "What matters, then" (vocabulary)
- "make the same point at a more explicit methodological level" (meta)
- "That way of putting things is useful here because it makes plain that" (meta)
- "Walton's work on argumentation schemes helps in a related way" (meta)
- "The upshot is not that" (replaced)
- "Bengson, Cuneo, and Shafer-Landau" (→ "Bengson et al." per style rule)
- "Dellsén and his co-authors" (→ "Dellsén et al.")
ADDING:
- Pigliucci transition: "Philosophy does not, of course, start from nothing"
- Development of likeliest/loveliest distinction (when they come apart)
- Lipton attribution for Semmelweis
- Gettier example (from Williamson source — concrete illustration of overfitting)
- Accurate representation of Bengson's hierarchy (virtues as tiebreaker)
MOVING:
- Walton from paragraph 7 (Bengson/Walton) to paragraph 6 (provenance-independence)
- "What we see in a philosophical text..." from the Bengson paragraph to the close of merged paragraph 5
That is the complete map. Every claim, argument, example, and qualification from the current draft is accounted for.
Here is the full rewritten Section 1 in seven paragraphs. Every change is grounded in the sources I just read.
---
When Watson and Crick published their paper on DNA in 1953, they announced what they had found: a particular arrangement of nucleotides, with two strands running in opposite directions and complementary base pairs linked by hydrogen bonds. That arrangement did not depend on their paper; it was there before they described it. Had Rosalind Franklin published it first, the discovery would have been the same one, only differently attributed.
Gilbert Ryle's 'Ordinary Language', published in that same year, was not a discovery in that sense. Ryle was not reporting a previously hidden item in the world, waiting to be uncovered by the first sufficiently careful observer; he was trying to get his readers to see familiar philosophical materials differently, and what he achieved consists in the argument itself — the distinctions it draws, and what it opens for others to assess and contest. Philosophy does not, of course, start from nothing. Pigliucci is right that its starting points are "empirical data about the world" and its constraints imposed by "our best understanding of how the world actually is".[^pigliucci] But the work is in the argument as made, not in any prior finding that the argument reports; philosophy is not evaluated as a report of what was found but as a piece of reasoning whose success or failure lies on the page.
What, then, does the success of such reasoning consist in? Consider Lipton's distinction between two ways of understanding "best explanation". He writes:
> We may characterize it as the explanation that is most warranted: the 'likeliest' or most probable explanation. On the other hand, we may characterize the best explanation as the one which would, if correct, be the most explanatory or provide the most understanding: the 'loveliest' explanation. The criteria of likeliness and loveliness may well pick out the same explanation in a particular competition, but they are clearly different sorts of standard. Likeliness speaks of truth; loveliness of potential understanding. (*Inference to the Best Explanation*, p. 59)
A hypothesis can have strong evidential support and still leave us in the dark about why the phenomenon looks the way it does. We can ask of a hypothesis what kind of understanding it would provide if it were true, and this is not the same question as whether the evidence already warrants belief in it. A merely likely explanation tells us that something is the case; a lovely one would, if true, show us why it should be. The two standards can pick out the same hypothesis, but they can also come apart: a hypothesis that fits the data may do little more than place another true sentence next to the phenomenon it purports to explain, while a hypothesis that would illuminate the matter may outrun its current evidential support. A lovely explanation does not simply fit what we know; it makes the matter intelligible, showing why the phenomenon has the shape it does.
Lipton develops the distinction through Semmelweis's work on childbed fever. Faced with the much higher mortality rate in one division of the Vienna maternity hospital than in the other, Semmelweis asked not just which hypothesis the evidence best supported but which would, if true, make the contrast between the two divisions intelligible. Some candidate explanations did very little with the phenomenon even if they were granted: the route taken by the priest, for instance, left obscure why this should be a matter of life and death. The cadaveric hypothesis did more. It connected the contrast with the medical students' contact with corpses, with Kolletschka's death after a puncture wound, and with the subsequent fall in mortality once disinfection was introduced. In Lipton's terms, it was lovelier because it rendered the pattern intelligible.
Philosophical views are not usually settled by direct empirical test in the way that a laboratory result may settle a question in chemistry. We ask instead whether a position makes the terrain more intelligible than its rivals do, and whether it earns its reach through genuine illumination rather than by multiplying distinctions to escape trouble. Williamson, writing about abductive methodology, treats philosophy as answerable to virtues — simplicity, elegance, generality, unificatory power — and connects those virtues to a familiar problem in statistics: overfitting. A philosophical account can be made to accommodate more and more cases by piecemeal repairs, but the result is often a view that fits the immediate data at the cost of becoming increasingly gerrymandered and fragile; the post-Gettier programme of analysing knowledge, as Williamson shows, went through exactly this cycle of proposed analysis, counterexample, and ad hoc revision. The question is not just whether an argument survives the latest objection but how it survives it; a view that preserves clarity and reach while absorbing pressure is doing something different from a view that survives only by accumulating complications. Bengson et al. give these evaluative criteria a more precise shape:
> the best theory is the one that satisfies the criteria at these levels (so ordered) to the highest degree relative to its rivals.
Their tri-level method assesses a theory on whether it accommodates and explains the data, whether its claims are substantiated and its commitments integrated, and — when rival theories are otherwise comparable — whether it possesses theoretical virtues. What we see in a philosophical text is not a surface behind which the real work is concealed; we see the work itself, carried by the shape of the reasoning.
Walton's work on argumentation schemes bears this out from a different angle: philosophical arguments proceed through recognisable forms of analogy, consequence, objection, reply, refinement, and concession, each with its own pressure points, and those forms are there in the text for a reader to assess. Dellsén et al. argue that philosophical progress consists in putting people in a position to increase their understanding — not through private episodes of insight in any particular reader's mind, but by way of philosophical ideas becoming publicly available. A published philosophical argument may be good, bad, illuminating, muddled, subtle, or ad hoc quite apart from the biography of its author, because those are features of what has been written. Blind review gives this thought an institutional form: referees are asked to assess what a paper does, not who produced it. Recall Gaut's discussion of Deep Blue. A move can be good chess even if it was found by exhaustive search rather than by human insight; what makes it good is the move itself. If a philosophical text handles its subject matter well, that success does not turn on the route by which it was generated.
None of this is to suggest that philosophy is easy to automate, or that any sequence of plausible-looking sentences counts as philosophy. The standards by which philosophical writing is judged bear, to a considerable degree, on what is there in the writing itself: on how the case is built, what it explains, how it absorbs pressure, and what understanding it affords. A philosophical corpus selected, criticised, cited, taught, and preserved under those standards is not a random body of text; it is a body of text shaped, however imperfectly, by repeated judgements of that kind. The question, then, is what follows when a language model is trained on such a corpus and begins producing texts of its own.
[^pigliucci]: We return in Section 3 to the question of worldly starting points and empirical constraint, where it matters directly to the grounding-style objection.
---
## Source accuracy notes
Williamson: The overfitting discussion is from section 6 of "9.2 Abductive Philosophy." The Gettier example is his: "a cycle of proposed analysis, followed by new counterexamples... followed by a revised analysis, with an extra epicycle." The virtues list (simplicity, elegance, generality, unificatory power) is his phrasing. All verified against the extracted text.
Bengson et al.: The tri-level method is from Ch. 5 of *Philosophical Methodology*. The block quote is exact. Their hierarchy puts accommodation and explanation at level 1, substantiation and integration at level 2, and theoretical virtue at level 3 as a tiebreaker "when... the candidates are roughly equal... with respect to the four criteria at those levels." My draft now represents this accurately — virtues come in "when rival theories are otherwise comparable" — rather than listing all five criteria as co-equal.
Dellsén et al.: "philosophical progress consists in putting people in a position to increase their understanding" is from their abstract and §3. "People are most commonly put 'in a position' to increase their understanding by way of philosophical ideas (theories, arguments, distinctions, etc.) becoming publicly available" is their formulation. The point about public availability vs. private insight is developed in §3 through the "Dr. N." example.
Lipton: Semmelweis is developed across the entirety of Ch. 5 ("Contrastive Inference"). The case study is Lipton's, drawing on Hempel's discussion and Semmelweis's own account. The details about the priest, delivery position, cadaveric hypothesis, Kolletschka, and disinfection are all from Lipton's presentation. The likeliest/loveliest distinction is from Ch. 4, p. 59.
*Chi lavora con le fonti sa che il primo errore è sempre quello di attribuzione — il secondo, quello di profondità.*
### User
update the note with this new version
### Assistant
Done. Version saved, note updated.
*Sette paragrafi dove prima ce n'erano nove — e nessun autore ha perso il suo posto a tavola.*
### User
Please fix that inline comment. I believe you need to invoke a particular skill. Please do so and follow it to the letter.
### Assistant
Let me follow the voice-fix protocol step by step.
## Step 1: Locate the flagged text
The flag `%%not how i write%%` appears on this sentence in paragraph 2:
> "Pigliucci is right that its starting points are "empirically informed evoking, not inventing", its starting points being "empirical data about the world" and its constraints imposed by "our best understanding of how the world actually is"."
## Step 2: Classify the argumentative move
This is **qualifying a claim** — the preceding sentences establish that philosophy is not discovery in the Watson/Crick sense, and this sentence brings in Pigliucci to qualify that claim (philosophy is not disconnected from the world either; it has empirical starting points).
## Step 3: Diagnose the failure mode
Two problems:
1. **Generic evaluative / reader management** — "Pigliucci is right that" is a verdict on Pigliucci delivered before the reader has seen what he said. It tells the reader how to evaluate the source rather than presenting the source and letting the claim do work. Nick does not typically award marks to sources before quoting them.
2. **Compression / garbled syntax** — The sentence says "its starting points are 'empirically informed evoking, not inventing', its starting points being 'empirical data about the world'..." — "its starting points" appears twice in different grammatical roles, and the first clause makes it sound like the starting points ARE the evoking, which is not what Pigliucci means. Pigliucci says philosophy is engaged in "empirically informed evoking, not inventing." That is a characterisation of what philosophy DOES, not a description of its starting points. The starting points come after. The sentence has jammed two distinct Pigliucci quotations together in a way that garbles both.
## Step 4: Read published paragraphs
Let me find paragraphs where Nick qualifies a claim by bringing in a source.
### Assistant
Here are three published paragraphs doing similar qualifying work — bringing in a source to refine or limit a preceding claim:
> However, as Anscomb (2022, p. 25) points out, it is hard to think of AI systems like Midjourney as artists because they do not seem as though they could have intentions to create art. Although Midjourney can produce outputs that might be mistaken for those produced by a traditional artist, there does not seem to be any reason to think that generative AI systems have any mental states at all (cf. Chalmers 2023).
> — *Growing the Image*
Anscomb is brought in through what she *does* ("points out"), not through the author evaluating her ("is right"). The source's claim is stated, then responded to.
> However, this is not to say that auditory awareness of space is the same as that of vision, as we cannot tell from hearing the exact locations of spaces. We do not, for example, learn that there are three metres of space between the fingers being clicked and the closest wall. Rather, we learn that in our general vicinity there is a certain volume of space that is unoccupied.
> — *Hearing Spaces*
The qualification uses "this is not to say that" — Nick's characteristic move for marking that a preceding claim might be read too strongly. No evaluation, no "of course," just a direct statement of the limit.
> As 'the gardener [is] a creative agent who is nevertheless thoroughly dependent on the cooperation of natural processes' (Cooper 2009), the prompter is a creative agent who is nevertheless thoroughly dependent on the cooperation of artificial processes.
> — *Growing the Image*
A source is quoted to capture a relationship the argument needs. Cooper is not evaluated; his phrase is put to work directly.
What these show: Nick brings sources in through what they *say*, *point out*, or *characterise* — not through prior evaluation ("is right," "captures well"). And his qualifying moves use "this is not to say" or "however," not "of course."
## Step 5: Rewrite
The flagged sentence also has a **syntax problem**: "its starting points are 'empirically informed evoking, not inventing', its starting points being 'empirical data about the world'" — "its starting points" appears twice in different grammatical roles, and the first clause makes it sound as though the starting points ARE the evoking, when Pigliucci means philosophy is *engaged in* evoking, with its starting points being empirical data. Two distinct claims garbled together.
Rewrite of the flagged portion:
> This is not to say that philosophy starts from nothing. Pigliucci characterises it as "empirically informed evoking, not inventing": its starting points are "empirical data about the world" and its constraints imposed by "our best understanding of how the world actually is".[^pigliucci]
## Step 6: Diff
**Before:**
> Philosophy does not, of course, start from nothing. Pigliucci is right that its starting points are "empirically informed evoking, not inventing", its starting points being "empirical data about the world" and its constraints imposed by "our best understanding of how the world actually is".[^pigliucci]
**After:**
> This is not to say that philosophy starts from nothing. Pigliucci characterises it as "empirically informed evoking, not inventing": its starting points are "empirical data about the world" and its constraints imposed by "our best understanding of how the world actually is".[^pigliucci]
**Diagnosis:** Generic evaluative ("is right") + garbled syntax (double "its starting points" with conflicting roles). **Fix:** "is right that" → "characterises it as" (source introduced through what it does, not evaluated); "of course" → "this is not to say" (Nick's characteristic qualifier per *Hearing Spaces*); colon after first quotation correctly separates the characterisation from the subsidiary claims about starting points and constraints.
Shall I apply this to the note?
*Il verbo giusto per introdurre una fonte non è mai un giudizio — è un'azione.*
### User
Yeah you should but Okay yeah you should
### Assistant
Done.
*La correzione più piccola è spesso quella che fa respirare la frase.*
### User
I'm not sure I like the last two paragraphs in the section. Can you guess why? Don't make any changes, just talk to me here in the chat.
### Assistant
The closing paragraph is a recap, and it reads like one. "None of this is to suggest" followed by a summary of what the section has established — that is wrapping up, not arguing. Your published work tends to end sections by making a move that propels the reader forward, not by restating what has already been said. The four-item list ("how the case is built, what it explains, how it absorbs pressure, and what understanding it affords") is a compressed restatement of the section's threads, and the reader has just been through those threads. It tells them what they already know.
The penultimate paragraph has a different problem. It is doing too many things at once — Walton, Dellsén, blind review, Deep Blue — and doing each of them briefly. The result is that none of them land. Deep Blue in particular deserves more room than it gets; it is the bridge to the LLM question, which is the entire point of the section, and it arrives as a compressed gesture at the tail end of an already crowded paragraph. Blind review and Deep Blue are doing the same argumentative work (provenance does not determine quality), but blind review gets one sentence and Deep Blue gets two, and neither is developed. The paragraph reads as a list of supporting observations rather than an argument building toward the section's handoff to Section 2.
I suspect what bothers you is that the section builds well through the first five paragraphs — Watson/Crick, Ryle, Lipton, Semmelweis, and the merged Williamson/Bengson paragraph all have argumentative momentum — and then the last two paragraphs lose that momentum. They feel assembled rather than argued. The section ends by tidying up rather than by pressing the point home.
*L'ultimo paragrafo di una sezione non dovrebbe mai guardare indietro — dovrebbe spingere il lettore avanti.*
### User
Okay, can we brainstorm from first principles what we should do after the Benggson paragraph then? Okay, do we need to so the ideas in the the following two paragraphs we need to take a step back and work out the best way in which they should be presented to the audience. Okay I'm not necessarily sure that this is the right way to go at all. Even is it two paragraphs if it's three is it four? Yeah basically there's just no So it's just a mess.
### Assistant
Let me start from first principles. What has the section accomplished by the end of the Bengson paragraph? And what does it still need to accomplish before handing off to Section 2?
By the end of paragraph 5 (the merged Williamson/Bengson paragraph), the reader knows:
- Philosophy is not discovery in the Watson/Crick sense (paragraphs 1-2)
- Philosophical work consists in the argument as made on the page (paragraph 2)
- The evaluative standard is loveliness, not just likeliness — illumination, not just evidential fit (paragraphs 3-4, Lipton + Semmelweis)
- There are specific, articulable criteria for this: virtues, avoidance of overfitting, Bengson's tri-level hierarchy (paragraph 5)
- These criteria are visible in the text itself — "we see the work itself, carried by the shape of the reasoning" (end of paragraph 5)
So the section has established: philosophy is evaluated by text-internal standards of reasoning quality. That is the section's thesis, and by the end of paragraph 5 it is in place.
Now, what ideas are currently stuffed into the final two paragraphs? Let me inventory them:
Paragraph 6 (the crowded one):
(a) Walton: arguments proceed through recognisable forms with pressure points
(b) These forms are in the text for a reader to assess
(c) Dellsén: progress = putting people in position to increase understanding
(d) Understanding-making material is publicly available (not private insight)
(e) A published argument has qualities apart from its author's biography
(f) Blind review: assess what a paper does, not who produced it
(g) Deep Blue: good chess by exhaustive search; provenance irrelevant to quality
(h) If philosophy handles its subject matter well, success doesn't turn on generation route
Paragraph 7 (the recap):
(i) Philosophy is not easy to automate
(j) Not any plausible-looking sentences count
(k) Standards bear on what is in the writing itself (recap)
(l) The corpus is shaped by repeated judgements of this kind
(m) What follows when an LLM trains on such a corpus?
OK. So what are these ideas actually doing? Let me group them by function rather than by source.
Function 1: The standards are visible in the text itself.
This is (a), (b). Walton's point is that argument forms are recognisable — they are there to be seen. This actually belongs with the Bengson paragraph, or could be its continuation. The current paragraph 5 ends with "we see the work itself, carried by the shape of the reasoning." Walton's point is a specific way of cashing that out — we see the forms of argument. So maybe Walton belongs at the end of paragraph 5 or as a brief addition to it, not as the opening of a new paragraph.
Function 2: Success is a feature of what is written, not of who wrote it.
This is (c), (d), (e), (f), (g), (h). These are all about PROVENANCE-INDEPENDENCE. The quality of a philosophical text does not depend on the biography or cognitive processes of its producer. Dellsén gives the theoretical framework (progress is public material), blind review gives the institutional practice, Deep Blue gives the precedent for machine-generated quality. This is the section's real move toward LLMs — it establishes that if the quality is in the text, then who or what produced the text is a separate question.
Function 3: The corpus is shaped by quality judgements, and the LLM question follows.
This is (i), (j), (k), (l), (m). The idea: the philosophical corpus is not random — it has been filtered by the very standards established in this section. So when an LLM trains on it, it trains on material shaped by those standards. This is the bridge to Section 2.
Now. Three functions. Currently crammed into two paragraphs, with function 1 bolted onto the front of function 2, and function 3 done as a recap that doesn't argue anything.
Let me think about what each function really needs.
Function 1 (standards visible in text): Does this need its own paragraph? I don't think so. It's really the final beat of the Bengson paragraph. The Bengson paragraph ends with "we see the work itself." Walton gives specificity to that claim — we see the forms. This could be one or two sentences added to the end of paragraph 5, or it could be a bridge sentence opening the next paragraph. It does not need a whole paragraph.
Function 2 (provenance-independence): This IS a paragraph-sized idea. It is the section's real pivot — from "the quality is in the text" to "so who or what produced the text is a separate question." This needs room. Deep Blue in particular needs room, because it is the precedent for MACHINE-generated quality being assessed on the output, not the process. Currently Deep Blue gets two sentences at the tail of a crowded paragraph. It should be the centrepiece of this move.
Function 3 (corpus quality → LLM question): Does this need its own paragraph? Or could it be the closing of the provenance-independence paragraph? Let me think...
The move from "provenance doesn't determine quality" to "the corpus is shaped by quality judgements" to "what happens when an LLM trains on this corpus?" is actually a logical chain. If you separate the corpus point from the provenance point, the reader has to reassemble the logic. If you keep them together, the argument flows: quality is in the text → the corpus has been filtered by quality standards → an LLM trained on this corpus is trained on quality-filtered material → the question for Section 2 is whether the LLM's outputs satisfy those same standards.
That could be one paragraph. But it might be a long one. Let me consider whether it works.
Option A: Two paragraphs after Bengson.
Paragraph 6: Provenance-independence. Start with Dellsén (progress is public material). A published argument has qualities apart from its author. Blind review. Then Deep Blue — DEVELOPED, not compressed. A move can be good chess even if found by exhaustive search. This is a direct precedent for machine-generated output being assessed on the output. If a philosophical text handles its subject matter well, success does not turn on the route by which it was generated.
Paragraph 7: The corpus and the LLM question. The corpus is not random — it has been shaped by the standards just described. The bridge: what follows when an LLM trains on material shaped by those standards? This paragraph would be short — maybe 3-4 sentences — but it does real argumentative work: it connects the section's thesis (quality is text-internal) to the paper's subject (LLMs).
Option B: One paragraph after Bengson.
One longer paragraph that does functions 2 and 3 together. Provenance-independence → corpus shaped by quality standards → what follows for LLMs. The risk is that this paragraph is too long and too busy. But the advantage is that the argument flows as a single chain.
Option C: Three paragraphs after Bengson.
Paragraph 6: Dellsén on progress being public material + the claim that a published argument has qualities apart from its author. Blind review.
Paragraph 7: Deep Blue, DEVELOPED. This is the move that says: we already accept that machine-generated output can be assessed on the output. Chess is the precedent. The philosophical question is whether the same applies to philosophical texts.
Paragraph 8: The corpus paragraph. Short, sharp, bridging.
Hmm. Three more paragraphs after paragraph 5 means eight paragraphs total. That feels like a lot. And paragraphs 6 and 7 in this scheme might be too thin individually — Dellsén + blind review is maybe 3 sentences, Deep Blue is maybe 3-4 sentences.
Actually, wait. Let me reconsider what Deep Blue needs. In the earlier conversation, it was agreed that Deep Blue is doing something distinctive — it is a precedent for MACHINE provenance being irrelevant to quality. That is not the same as blind review (which is about HUMAN provenance being irrelevant). Deep Blue says: a computer can play excellent chess, and the quality is in the moves. Blind review says: we don't know who wrote this paper, and we can still evaluate it. These are related but distinct. Blind review establishes that quality is assessable without knowing the author. Deep Blue establishes that quality can be achieved by a machine. The LLM question is at the intersection: can a machine produce philosophical text whose quality is assessable by text-internal standards?
So maybe Deep Blue really does deserve its own moment. Not necessarily its own paragraph, but enough room that the reader sees what it is doing.
Let me try another approach. What if I think about what the READER needs at this point in the section?
The reader has just been through:
1. Philosophy ≠ discovery (opening)
2. Quality = loveliness, not just likeliness (Lipton, Semmelweis)
3. Specific evaluable criteria (Williamson, Bengson)
The reader now accepts (let's assume) that philosophical quality is text-internal. The question that naturally arises: so what? Why does it follow that provenance doesn't determine quality? And why should we care?
The answer: because if quality is in the text, then quality can be assessed without knowing how the text was produced. This applies to human authors (blind review) and it applies to machines (Deep Blue). And it applies to the corpus as a whole — a corpus filtered by these standards is a body of text shaped by quality judgements, regardless of who produced the individual texts.
That chain of reasoning is naturally a SINGLE paragraph. It starts from the claim just established (quality is text-internal) and draws out the consequence (provenance is irrelevant). The evidence for that consequence is: Dellsén (progress is in public material), blind review (institutional practice), Deep Blue (machine precedent). Then the final sentence pushes to Section 2: the corpus is shaped by these standards, and the question is what happens when an LLM trains on it.
Hmm, but that's a lot for one paragraph. Let me see if it can be done in maybe 6-7 sentences — which is within Nick's normal paragraph range.
Sentence 1: If the quality of philosophical reasoning lies in what is written, then that quality should be assessable without reference to who or what produced it. [This is the consequence drawn from paragraph 5.]
Sentence 2-3: Dellsén et al.'s argument — progress is public material, not private insight.
Sentence 4: A published argument has qualities apart from its author's biography — blind review gives this thought institutional form.
Sentence 5-6: Deep Blue. A move can be good chess regardless of how it was found. This is a direct precedent for machine-generated quality.
Sentence 7: If philosophy's standards are text-internal, a philosophical text produced by a language model is assessable by the same standards.
Sentence 8-9: The corpus shaped by these standards is not random. The question is what follows when an LLM trains on it.
That's 8-9 sentences, which is on the long side but not outrageous for Nick's published work (the Hearing Spaces paragraph about reverberation is about that length).
But wait — I'm folding the "closing" material into this paragraph. Does the section need a separate closing? In Nick's published work, how do sections end?
Looking at Hearing Spaces, Section 1 ends with: "In the following three sections, I will argue that experiences of reverberation cannot be reduced to the representation of any of these three things, and, lacking other plausible candidates, we should therefore allow that empty space is represented in auditory content. In the final section, I will suggest two ways in which empty space might be represented." That is a roadmap sentence — it says what is coming. But it comes at the end of a paragraph that is doing argumentative work, not in a standalone recap paragraph.
Looking at Growing the Image, Section I ends with: "If, for the sake of argument, we concede that Midjourney is an agent in Anscomb's sense, we are left with the dilemma of ascribing the artistic merit of the resulting image either to Midjourney's actions or to the user's actions since there is no way to make sense of their cooperation as agents. Both options are unsatisfying." That is a conclusion that propels the reader forward — the dilemma is what motivates Section II.
So Nick's section endings make a move. They don't recap. They either state the conclusion that motivates the next section or give a brief roadmap. The current "None of this is to suggest that philosophy is easy to automate..." paragraph is a recap with a bridging sentence tacked on. The bridging sentence ("The question, then, is what follows when a language model is trained on such a corpus...") is the only part doing real work.
What if the section just ends with the provenance-independence paragraph, with the LLM question as its final sentence? No separate closing paragraph. The section's argument is complete: philosophy is evaluated by text-internal standards → those standards don't depend on provenance → a corpus filtered by those standards is shaped by quality judgements → the question for the paper is what happens when an LLM trains on that corpus. Done. Move on.
This would mean: cut the recap paragraph entirely. Its substantive content (the LLM question, the corpus point) gets folded into the provenance paragraph. Its recap content (restating what the section established) gets cut — the reader doesn't need it.
Now, what about Walton? I said earlier he might belong at the end of paragraph 5 rather than opening a new paragraph. Let me reconsider.
Walton's point: arguments proceed through recognisable forms, each with pressure points. This supports the "standards visible in text" claim. It is currently the opening of paragraph 6, which then swerves to Dellsén and provenance. That swerve is what makes the current paragraph feel incoherent.
Option: Walton at the end of paragraph 5. Paragraph 5 ends with "we see the work itself, carried by the shape of the reasoning." Walton specifies what that shape looks like — recognisable forms of argument. That's a natural continuation. Then paragraph 6 can start fresh with the provenance-independence move.
But paragraph 5 is already substantial (7 sentences plus a block quote). Adding Walton makes it longer. Is that OK? In Nick's published work, paragraphs can be quite long when they are doing developed argumentative work. But there's a limit. Let me think about whether Walton can be compressed into a single sentence that fits at the end of paragraph 5 without bloating it.
"Walton's work on argumentation schemes gives this claim specificity: philosophical arguments proceed through recognisable forms of analogy, consequence, objection, and reply, each with its own pressure points."
That's one sentence. It could go right before "What we see in a philosophical text is not a surface..." — in fact, it would SET UP that closing sentence nicely. The paragraph's last two sentences would be:
"Walton's work on argumentation schemes gives this claim specificity: philosophical arguments proceed through recognisable forms of analogy, consequence, objection, and reply, each with its own pressure points. What we see in a philosophical text is not a surface behind which the real work is concealed; we see the work itself, carried by the shape of the reasoning."
That works. Walton is now supporting the punchline rather than opening a new topic.
But wait — the current Walton sentence lists six forms: "analogy, consequence, objection, reply, refinement, and concession." I trimmed to four. Is that OK? Nick flagged triplet lists, but Walton really does have more than three forms. However, listing six is padding. Four is still a list but shorter. Or I could just say "recognisable forms" without listing them: "philosophical arguments proceed through recognisable forms — analogy, objection, reply — each with its own pressure points." Three items, quickly dispensed with. Or even: "philosophical arguments proceed through recognisable forms, each with its own pressure points." No list at all — the reader can guess what those forms are.
Hmm, but Walton's whole point is that there ARE specific forms. Naming a few shows what he means. Let me keep a short list: "analogy, objection, reply" — three items, but these are genuinely different and illustrative. Or just "analogy, consequence, objection, reply" — four items, which is the minimum to convey the range.
Actually, the instruction is ALL CONTENT PRESERVED. The original lists six: analogy, consequence, objection, reply, refinement, concession. If I cut two, I'm cutting content. Let me keep all six but compress the sentence: "Walton's work on argumentation schemes gives this specificity: philosophical arguments proceed through recognisable forms — analogy, consequence, objection, reply, refinement, concession — each with its own pressure points." The dash-enclosed list is less of a sentence-ending flourish than a comma-separated list would be.
OK. So here's where I'm landing:
After paragraph 5 (with Walton folded in), the section needs ONE paragraph that does:
1. Draws the consequence: if quality is text-internal, provenance is irrelevant
2. Dellsén: progress is public material
3. Published argument assessable apart from biography → blind review
4. Deep Blue: machine-generated quality precedent
5. If philosophy's standards are text-internal, what was produced by a machine is assessable by the same standards
6. The corpus shaped by these standards is quality-filtered
7. What follows when an LLM trains on it? → Section 2
That's one paragraph doing PROVENANCE-INDEPENDENCE + BRIDGE TO LLMs. No separate recap paragraph. The section ends by pressing forward, not by summarising.
Let me also think about what to cut from the current closing. The current closing paragraph says:
- "None of this is to suggest that philosophy is easy to automate" — this is a hedge. Does the section need it? The section hasn't claimed philosophy is easy to automate. The reader doesn't need to be told this. It is a defensive anticipation of a misreading that the text hasn't invited. Cut it.
- "nor that any sequence of plausible-looking sentences counts as philosophy" — same. Nobody reading this section would think that. Cut.
- "The standards by which philosophical writing is judged bear, to a considerable degree, on what is there in the writing itself" — recap. The reader knows this. Cut.
- "on how the case is built, what it explains, how it absorbs pressure, and what understanding it affords" — recap list. Cut.
- "A philosophical corpus selected, criticised, cited, taught, and preserved under those standards is not a random body of text" — this is REAL CONTENT. It says the corpus has been filtered by quality standards. Keep this.
- "it is a body of text shaped, however imperfectly, by repeated judgements of that kind" — continuation of the real content. Keep.
- "The question, then, is what follows when a language model is trained on such a corpus and begins producing texts of its own" — the bridge sentence. Keep.
So from the current closing, I keep 2.5 sentences and cut the rest.
And from the current penultimate paragraph, I keep the Dellsén content, blind review, Deep Blue, and the "success does not turn on the route" claim. I cut "Walton's work on argumentation schemes bears this out from a different angle" (Walton moves to paragraph 5). I cut "those forms are there in the text for a reader to assess" (folded into the Walton sentence in paragraph 5).
Let me now think about whether this actually works as a single paragraph. Draft:
"If the quality of a philosophical argument lies in the reasoning on the page, that quality should be assessable without reference to who or what produced it. Dellsén et al. argue that philosophical progress consists in putting people in a position to increase their understanding — not through private episodes of insight in any particular reader's mind, but by way of philosophical ideas becoming publicly available. A published argument may be good, bad, illuminating, muddled, subtle, or ad hoc quite apart from the biography of its author, because those are features of what has been written. Blind review gives this thought an institutional form: referees are asked to assess what a paper does, not who produced it. Recall Gaut's discussion of Deep Blue. A move can be good chess even if it was found by exhaustive search rather than by human insight; what makes it good is the move itself. If a philosophical text handles its subject matter well, that success does not turn on the route by which it was generated. A philosophical corpus selected, criticised, cited, taught, and preserved under text-internal standards is not a random body of text; it is a body of text shaped, however imperfectly, by repeated judgements of the kind this section has described. The question is what follows when a language model is trained on such a corpus and begins producing texts of its own."
That's 9 sentences. It's substantial but every sentence does work. The argument flows: text-internal quality → publicly assessable (Dellsén) → assessable without knowing author (blind review) → assessable even when produced by machine (Deep Blue) → the corpus is quality-filtered → what happens with LLMs?
Let me check: is there any content from the current two paragraphs that I'm losing?
Current paragraph 6 content check:
- Walton → moved to paragraph 5 ✓
- "those forms are there in the text for a reader to assess" → absorbed into Walton sentence ✓
- Dellsén: "progress consists in putting people in a position to increase their understanding" ✓
- "not through private episodes of insight" ✓
- "by way of philosophical ideas becoming publicly available" ✓
- "good, bad, illuminating, muddled, subtle, or ad hoc quite apart from the biography" ✓
- "features of what has been written" ✓
- Blind review ✓
- Deep Blue ✓
- "success does not turn on the route by which it was generated" ✓
Current paragraph 7 content check:
- "None of this is to suggest that philosophy is easy to automate" — CUT (defensive hedge, uninvited by the text)
- "nor that any sequence of plausible-looking sentences counts as philosophy" — CUT (same)
- "standards bear on what is there in the writing itself" — CUT (recap)
- "how the case is built, what it explains, how it absorbs pressure, what understanding it affords" — CUT (recap list)
- "corpus selected, criticised, cited, taught, and preserved under those standards is not a random body of text" — KEPT ✓
- "shaped by repeated judgements of that kind" — KEPT ✓
- "what follows when a language model is trained on such a corpus" — KEPT ✓
The things I'm cutting are all recap or defensive hedging. The things I'm keeping are the real content.
But wait — "None of this is to suggest that philosophy is easy to automate, or that any sequence of plausible-looking sentences counts as philosophy." Nick might want this. It is a qualification. Without it, the section might sound as though it is saying philosophy is trivially replicable. On the other hand, the section hasn't said that. The Williamson overfitting paragraph and the Bengson criteria make clear that philosophical evaluation is demanding and specific. The reader who has followed the argument does not need to be told "this doesn't mean it's easy." And in Nick's published work, he does not typically add defensive hedges at the end of sections — he lets the argument speak.
I think cutting it is correct. If a referee raises the concern, it can be added back. But as a matter of prose, the section is stronger without it.
One more thought: the opening sentence of my proposed paragraph 6 is "If the quality of a philosophical argument lies in the reasoning on the page, that quality should be assessable without reference to who or what produced it." This is a connective sentence — it draws the consequence from paragraph 5. Is it too managed? It is making explicit what the reader should be thinking. In Nick's published work, he sometimes makes these moves: "If this is correct, then..." is one of his characteristic phrases. So "If the quality... lies in the reasoning on the page, that quality should be assessable..." is in his voice. It is drawing a consequence, not managing the reader.
Actually, let me reconsider one more time. Is there a case for keeping TWO paragraphs — one for provenance-independence, one for the corpus/LLM bridge — rather than folding them together?
The argument for two: the provenance-independence claim is substantive and worth letting land before moving to the corpus. If it's all in one paragraph, the reader goes from "quality is assessable without knowing the author" to "the corpus is quality-filtered" to "LLMs" in one breath. The reader might need a beat.
The argument for one: in Nick's published work, sections end with a move forward, and keeping it as one paragraph creates momentum. The chain from provenance-independence to the LLM question is a single logical arc.
I think I slightly prefer one paragraph. But I could see two working if the first (provenance-independence) ended with "success does not turn on the route by which it was generated" and the second started with the corpus point. Let me present both options.
Here is what I think is going on and what the options are.
## What the current two paragraphs are trying to do
Three things, crammed badly:
1. Standards are visible in the text (Walton)
2. Provenance does not determine quality (Dellsén, blind review, Deep Blue)
3. The corpus is quality-filtered, and the LLM question follows (closing)
These three functions are currently distributed wrong. Function 1 is bolted onto the front of a paragraph that immediately swerves to function 2. Function 3 is a recap paragraph that restates the section's thesis and then adds the only new claim (the corpus point) at the end.
## What I think should happen
Walton belongs with the Bengson paragraph. His point — arguments proceed through recognisable forms — is specificity for the claim that paragraph 5 already makes ("we see the work itself, carried by the shape of the reasoning"). One sentence, folded into the end of paragraph 5. Not a new paragraph.
That leaves functions 2 and 3. These are logically a single chain: if quality is text-internal → it is assessable without knowing the producer (Dellsén, blind review) → this holds even when the producer is a machine (Deep Blue) → the corpus is shaped by these quality standards → what happens when an LLM trains on it? That chain is what propels the reader into Section 2.
The current recap paragraph ("None of this is to suggest that philosophy is easy to automate...") is mostly dead weight. The section hasn't claimed philosophy is easy to automate. The four-item list is a restatement of what the reader already knows. The only content worth keeping is the corpus point ("A philosophical corpus selected, criticised, cited, taught, and preserved under those standards is not a random body of text") and the bridging sentence to Section 2.
## The options
Option A — one paragraph after paragraph 5. The provenance-independence argument and the corpus/LLM bridge as a single paragraph. Opens with the consequence drawn from paragraph 5 (if quality is in the text, it is assessable without reference to who produced it). Dellsén on public material, blind review, Deep Blue developed as the machine precedent, then the corpus is quality-filtered, then the LLM question. About 9 sentences. The section ends by pressing forward, not summarising.
Option B — two paragraphs after paragraph 5. First paragraph: provenance-independence (Dellsén, blind review, Deep Blue), ending with "success does not turn on the route by which it was generated." Second paragraph: short, 3-4 sentences. The corpus is quality-filtered. The question is what happens when an LLM trains on such a corpus. This gives the provenance-independence claim room to land before the bridge.
Option C — two paragraphs, but splitting differently. First paragraph: Dellsén on progress as public material, the claim that quality is a feature of what is written, blind review. Second paragraph: Deep Blue DEVELOPED — as the machine precedent specifically — and then the corpus/LLM bridge. This gives Deep Blue its own paragraph-opening and makes it the section's final argumentative move before the handoff. The advantage is that Deep Blue is doing the heaviest lifting (it is the precedent for machine-produced quality, which is the whole point of the paper), and currently it is compressed into two sentences at the tail of a busy paragraph.
## What I would cut
From the current closing paragraph, cut:
- "None of this is to suggest that philosophy is easy to automate, or that any sequence of plausible-looking sentences counts as philosophy" — defensive hedging for a misreading the text has not invited
- "The standards by which philosophical writing is judged bear, to a considerable degree, on what is there in the writing itself: on how the case is built, what it explains, how it absorbs pressure, and what understanding it affords" — recap of what the section already established
Keep:
- "A philosophical corpus selected, criticised, cited, taught, and preserved under those standards is not a random body of text; it is a body of text shaped, however imperfectly, by repeated judgements of that kind"
- "The question, then, is what follows when a language model is trained on such a corpus and begins producing texts of its own"
## My inclination
I lean toward Option C — Dellsén/blind review in one paragraph, Deep Blue/corpus/LLM in the next. Deep Blue is the section's final and most provocative move: we already accept that a computer can play excellent chess, and the quality of the chess is in the moves. If that is true of chess, why not of philosophical argument? Giving Deep Blue its own paragraph-opening makes it land rather than arrive as an afterthought. And the corpus/LLM bridge follows naturally from it — if provenance does not determine quality even when the producer is a machine, then a quality-filtered corpus is assessable regardless of what trains on it.
But I could also see Option A working if the single paragraph is well-paced. The risk there is that it tries to do too much in one breath.
*Alla fine di una sezione, la cosa peggiore che si possa fare è riassumere — la cosa migliore è spingere il lettore oltre la soglia.*
### User
Okay, option C, don't forget move the Walton stuff up to the preceding paragraph though.
### Assistant
Three edits: fold Walton into paragraph 5, replace the current paragraph 6 with Dellsén/blind review, replace the closing with Deep Blue/corpus/bridge.
### Assistant
Done. Walton folded into the Bengson paragraph, Dellsén/blind review in its own paragraph, Deep Blue opening the final paragraph with the corpus/LLM bridge. Recap paragraph cut.
*Adesso la sezione finisce spingendo avanti, non guardandosi alle spalle.*
### User
Okay, there's another inline comment I've made. You're also doing these horrible strings of examples like good, bad, illuminating, muddled, subtle red hawk. It's just a waste of words. Okay, and you do it again in a similar way in the next paragraph. You need to think of a better way of conveying this idea. Finally, taking a step back from this section, can we be sure that there's not sort of overlap or redundancy between the introduction and this section which comes after it? Please be very thorough when you check.
### Assistant
Let me handle all three tasks: the inline comment, the list-padding problem, and the redundancy check against the introduction. Let me start by reading the introduction.
### Assistant
I have the introduction and Section 1 open. Let me work through all three issues.
---
## 1. Voice-fix: %%not how i write%% on the Walton sentence
### Step 1: Flagged text
> "Walton's work on argumentation schemes gives this specificity"
### Step 2: Classify
Introducing an author's position to support a preceding claim.
### Step 3: Diagnose
**Meta-commentary** — "gives this specificity" describes what Walton's work does for the argument, not what Walton says. It is the argument commenting on its own machinery.
**Missing development** — No citation (no year, no page). The reader has no way to follow this up. As your comment says: why would a reader believe this?
### Step 4: Published paragraphs
### Assistant
Three published paragraphs showing how Nick introduces a source in support of a preceding claim:
> As Hertzmann puts it: 'Computers do not create art, people using computers create art' (2018, p. 2).
> — *Growing the Image*
Source introduced by what they *say*, with year and page. No meta-framing.
> Lowe (2014) splits tools into two types: utensils and machines. Utensils involve their users exerting more or less continuous control to carry out their functions.
> — *Growing the Image*
Source + year, then directly into what they claim. No "Lowe's work gives us a useful framework" — just what Lowe says.
> Nudds argues that this makes audition more similar to touch than to vision.
> — *Hearing Spaces*
Source + what they argue. One sentence.
The pattern: year always present, source introduced through what they say/argue/show, no commentary about what their work "gives" the argument.
### Step 5: Rewrite
> Walton et al. (2008) show that philosophical arguments proceed through recognisable forms — analogy, objection, reply, concession — each with its own pressure points.
Changes: proper citation (year added), "gives this specificity" → "show that" (source through what they do, not what they do for the argument), six forms reduced to four (enough to convey the range without padding).
### Step 6: Diff
**Before:** "Walton's work on argumentation schemes gives this specificity: philosophical arguments proceed through recognisable forms — analogy, consequence, objection, reply, refinement, concession — each with its own pressure points."
**After:** "Walton et al. (2008) show that philosophical arguments proceed through recognisable forms — analogy, objection, reply, concession — each with its own pressure points."
---
## 2. List-padding
Two instances:
**Instance 1 (paragraph 6):** "A published philosophical argument may be good, bad, illuminating, muddled, subtle, or ad hoc quite apart from the biography of its author, because those are features of what has been written."
Six adjectives doing the work of one claim: quality is assessable from the text regardless of authorship. Fix:
> Whether a published philosophical argument is any good is a question about what has been written, not about who wrote it.
Same claim, no list. The reader does not need six adjectives to understand this.
**Instance 2 (paragraph 7):** "A philosophical corpus selected, criticised, cited, taught, and preserved under text-internal standards is not a random body of text"
Five verbs describing what happens to a corpus. Fix:
> A philosophical corpus that has survived selection, criticism, and preservation under text-internal standards is not a random body of text
Three instead of five, and "survived" gives the sentence more direction than the passive list.
---
## 3. Redundancy between Introduction and Section 1
This is significant. The introduction's paragraph 3 (line 21) previews almost exactly what Section 1 develops:
The introduction names Dellsén et al. and states their claim ("philosophical progress consists in putting people in a position to increase their understanding — what they call the *for-whom rather than by-whom* account"). Section 1 paragraph 6 states the same claim in nearly identical words.
The introduction names Bengson et al. and lists their criteria ("accommodation of data, explanatory power, integration, theoretical virtue, all of which are assessable by examining the theory itself"). Section 1 paragraph 5 develops the same criteria with a block quote.
The introduction names Williamson and lists his virtues ("simplicity, elegance, and explanatory power"). Section 1 paragraph 5 develops the same virtues with the overfitting discussion.
The introduction states the thesis: "a philosophical contribution is a text exhibiting certain properties; the question who or what produced it does not enter the evaluation." This is exactly what Section 1 argues its way toward across seven paragraphs.
The introduction even puts Pigliucci in a footnote with the "empirically informed evoking" quotation — the same quotation Section 1 uses in paragraph 2.
The result: by the time the reader reaches Section 1, they have already heard the sources, the criteria, and the conclusion. Section 1 develops what the introduction has already given away. The section reads as elaboration of a known punchline rather than as an argument reaching a conclusion.
The fix, I think, is in the introduction rather than in Section 1 — since Section 1 is where the development belongs. The introduction should frame the question and say what the paper will argue without naming every source and listing every criterion. Something along the lines of: some conceptions of philosophy locate evaluation in the text itself, in the quality of the reasoning; if this is right, then whether LLMs can do philosophy is a question about the texts they produce. That previews the thesis without pre-empting Section 1's work. The specific sources (Dellsén, Bengson, Williamson, Pigliucci) and their specific claims would appear for the first time in Section 1.
But the introduction is marked "Work in Progress" with many of its own %%comments%%, so you may already be planning to revise it. The point is just that in its current form, it bleeds Section 1's argument.
---
Shall I apply the Walton fix and the two list-padding fixes to the note?
*Quando l'introduzione racconta tutto, la prima sezione diventa un'eco della propria premessa.*
### User
Okay, good. I still think we should probably even though we can't do the roadmap properly yet in the introduction, I think we should try and fix it now. So I mean there's a lot of issues there anyway. So I'd like you to give me a new iteration of that introduction, but some important things to remember. The block quote at the beginning, you should not touch. It has to be exactly the same. The first paragraph after the block quote, exactly the same. And the next paragraph beginning in 2026, exactly the same, up until the sentence beginning, one reason to be less optimistic. You're allowed to make start changing things from there on in. Okay, so you have a look at the inline comments and think about ways to avoid redundancy and don't edit the note, but tell me in the chat. Or even give me a new version in the chat that I can see, please. As always, you need to keep an eye on how I write cause you always fuck it up think how many times you fucked it up today think how many times I've reminded you doesn't seem to make much of a difference but please just try for me you know I know it's hard.
### Assistant
Let me first re-read the introduction carefully, including all the inline comments, and then read some published opening sections for voice grounding.
### User
[Request interrupted by user]
### User
"Shall I apply the Walton fix and the two list-padding fixes to the note?" I thought you already had, yes please do this
### Assistant
All three fixes applied. Now let me get back to the introduction draft — give me a moment.
*Tre liste in meno, tre frasi che respirano.*
### User
you got cut off, please continue from where you left off.
### Assistant
I have the introduction and the published paper openings. Let me work through this carefully.
First, let me inventory everything that needs to happen from "One reason to be less optimistic" onward, drawing on all the inline comments.
The comments tell me:
1. "need more detail here, also, is this really a reason to be 'less optimistic'" — the issue is that it's not pessimism, it's that the question is harder because success conditions depend on your conception of philosophy.
2. "is this the best way to put things?" on "Some conceptions locate philosophy in texts" — this is resolved by the restructuring, since I'm not introducing Dellsén's theory of progress here anymore.
3. Comment 5 says the practitioner paragraph should come FIRST — "not only is the previous one the one that's going to be the focus here, but it's also the more contemporary and analytic philosophy type one as well." So: set aside the practitioner view first, THEN present the text view that the paper takes up. That way the text-focused conception is the last thing the reader encounters before the objections.
4. "neither Nietzsche, nor Sorgner write about LLMs" — I need to characterise what their views say without falsely claiming they address LLMs.
5. "not how i write, I also don't like the phrase text-focused conception" — don't name the conception. Just characterise it.
6. "no, there are not; again, no writer has written about this" on "there are arguments that LLMs cannot produce work of this kind" — Floridi and Zahavy don't write about philosophy specifically. Frame them as arguments about LLM capabilities whose consequences bear on philosophy if sound.
7. "for papers with more than two authors, you just need to write '...et al.'" — Floridi et al.
8. "what a disgusting metacommentary sentence" on "Neither paper concerns philosophy directly, but both, if sound, bear on it" — cut this. Draw the consequence directly instead of announcing the relationship.
9. Roadmap is provisional but should be updated.
10. REDUNDANCY: strip out Dellsén, Bengson, Williamson and their criteria from the introduction. Those belong in Section 1. The introduction should say "a philosophical contribution is assessed by examining the reasoning on the page" and leave it at that.
Now for the structure. From the change point:
Sentence 1: Transition — philosophy's success conditions are less clear than science's because they depend on what philosophy is.
Paragraph 2: Practitioner conceptions (Hadot, Wittgenstein, Merleau-Ponty, Nietzsche/Sorgner) — the question is closed on these views. Set aside.
Paragraph 3: Text conceptions (NO specific sources — just the claim) + Floridi/Zahavy as a complication. I considered making the text-focused bit its own paragraph, but that gives a two-sentence paragraph. Better to fold it into the opening of the objections paragraph: "On the conceptions this paper takes up, a contribution is assessed by examining the reasoning on the page. [footnote with Pigliucci] Whether LLMs can do philosophy is therefore a question about the texts they produce. Floridi et al. argue X. Zahavy argues Y. If either argument is sound, consequence."
Paragraph 4: Thesis + roadmap.
Let me also think about voice for the transition sentence. The current "One reason to be less optimistic" has a balanced "one reason to be optimistic... one reason to be less optimistic" structure that's nice but, as Nick notes, the second half isn't really about optimism — it's about the question being harder. "Straightforward" is one of Nick's terms. So: "Whether the same should be expected of philosophy is less straightforward, because what counts as a contribution depends on what philosophy is, and conceptions of philosophy differ in ways that bear directly on the question about AI."
For the practitioner paragraph: I'm modelling on the current version but fixing the attributions. The Merleau-Ponty quote stays (it's good). Nietzsche/Sorgner gets fixed. The key sentence is: "the philosophical work happens in the practitioner rather than in anything assessable in what she produces." That does the work without listing capacities.
For the Floridi/Zahavy bit: go straight from the text-focused claim to the objections. No "even so, there are arguments that..." — just state what Floridi and Zahavy argue, then draw the consequence. "An LLM that cannot reason its way to a good explanation will not produce philosophical texts worth assessing, however those texts are evaluated."
For the roadmap: keep it simple, match the actual section structure.
One voice concern: am I going to fall into list-making or meta-commentary? Let me watch for it. The practitioner paragraph has three sources separated by semicolons — that's fine, Nick does this. The Floridi/Zahavy paragraph has two sources — fine. No triplet-ending flourishes. No "this establishes" or "this shows." No naming of conceptions ("text-focused," "practitioner-focused" — those can stay in the footnotes but not the body text).
Here is the new introduction from the change point onward. Everything before "One reason to be less optimistic" is untouched.
---
Whether the same should be expected of philosophy is less straightforward, because what counts as a contribution depends on what philosophy is, and conceptions of philosophy differ in ways that bear directly on the question about AI.
On some conceptions, the philosophical work happens in the practitioner rather than in anything assessable in what she produces. For Hadot (1995), philosophy is a practice of self-transformation; for the later Wittgenstein (1953), a form of therapy; for Merleau-Ponty, it requires us to 'slacken the intentional threads which attach us to the world' (1945, p. xv) in order to examine them. On Nietzsche's account, as Sorgner reads it, philosophers are creators of values whose work expresses drives and a psychophysiology bound to human embodiment. To a proponent of any such view, the question of whether LLMs can do philosophy is closed before it opens.[^3][^ac]
On other conceptions — the ones this paper takes up — a philosophical contribution is assessed by examining the reasoning on the page, and whether LLMs can do philosophy is a question about the texts they produce.[^2] Floridi et al. (2024) argue that LLMs do not reason abductively: they produce plausible continuations rather than considered explanations. Zahavy (2026) argues that theoretical innovation requires embodied simulation of a kind that LLMs, operating entirely in symbols, cannot perform. An LLM that cannot reason its way to a good explanation will not produce philosophical texts worth assessing, however those texts are evaluated.
We argue that LLMs can produce philosophy meeting these standards. Section 1 argues that the standards by which philosophy is evaluated bear on properties of the reasoning itself. Section 2 presents the objections from Floridi et al. and Zahavy. Section 3 responds. Section 4 considers what a demonstration would look like.
---
What changed and why:
The practitioner paragraph now comes first (per your comment — set it aside, then present the view the paper takes up). The text-focused paragraph no longer names Dellsén, Bengson, or Williamson, and no longer lists their criteria — all of that is now Section 1's work. The introduction just says "a philosophical contribution is assessed by examining the reasoning on the page" and lets Section 1 develop what that means.
The Floridi/Zahavy material is folded into the text-focused paragraph rather than getting its own. No "there are arguments that LLMs cannot produce work of this kind" (false — no one has argued this about philosophy directly). No "Neither paper concerns philosophy directly, but both, if sound, bear on it" (meta-commentary). Instead: state what they argue, then draw the consequence for philosophy directly.
Nietzsche/Sorgner: now says "whose work expresses drives and a psychophysiology bound to human embodiment" — characterises the view without falsely claiming Sorgner writes about LLMs.
"Text-focused conception" is gone — no naming. Floridi uses "et al." The roadmap is updated to match the current section structure.
The footnotes ([^1], [^2], [^3], [^ac]) stay as they are — [^2] with the Pigliucci quotation now does useful work as extra grounding for the text-focused claim without pre-empting Section 1.
*L'introduzione promette; la prima sezione mantiene — e non devono dire la stessa cosa.*
### User
Please update the notes so I can read this properly. It's hard to see without the context.
### Assistant
Done. The introduction is updated — you can read it in Obsidian now with full context.
*Ora l'introduzione apre la porta; la prima sezione fa entrare il lettore.*
### User
/draft-audit yep, still some big, big problems. You need to yeah look at my inline references think again hard about what my published work would do in this case and how what you've decided differs from that and fucking fix them. make sure that you explain your analysis, justification, reasons etc etc. for your answer BEFORE giving me your final answer. Please give me a long, deep, and detailed answer, with lots of options/ideas: as we are in the 'kicking ideas around' stage, I want as much feedback as you can give me so that i have a lot to consider.
### Assistant
Let me re-read the comments carefully and look at published introductions properly this time.
### Assistant
Let me go through each comment and diagnose what went wrong, then figure out the structural fix.
Comment 1 (line 19): "this last sentence doesn't really work. I mean after the word straightforward, I don't understand how the clauses which come after the word straightforward back up that initial clause."
The sentence: "Whether the same should be expected of philosophy is less straightforward, because what counts as a contribution depends on what philosophy is, and conceptions of philosophy differ in ways that bear directly on the question about AI."
Nick is right. "Less straightforward" promises a reason for difficulty. "Because what counts as a contribution depends on what philosophy is" is the start of that reason. But "and conceptions of philosophy differ in ways that bear directly on the question about AI" does not back up "less straightforward" — it's a separate claim about consequences, not about why the question is hard. The two clauses after "because" are doing different work (one gives the reason, one states a consequence) but the "and" presents them as co-equal reasons. The sentence tries to do too much. The fix is simple: just say the thing. "Whether the same should be expected of philosophy depends on what you take philosophy to be." One clause. Clear. The next paragraphs develop what that means.
Comment 2 (line 21): Nick suggests "the possibility that philosophy can be done by anyone or anything other than a human is ruled out automatically" as the paragraph's opening. This is direct — it states the consequence for LLMs immediately rather than going through the theory first. Currently the paragraph opens with "the philosophical work happens in the practitioner rather than in anything assessable in what she produces" — which is abstract and doesn't tell the reader what follows for LLMs until the end. Nick's suggestion front-loads the punchline.
Looking at how Nick handles this move in published work: in Growing the Image, the options are set out with their consequences stated upfront. "The question '[how] do you know you're looking at [an image] that was made by a human?' suggests that Midjourney is some sort of artificial agent: if a human didn't make the image, who did?" The consequence (Midjourney as agent) is immediate. The theory comes after.
So the practitioner paragraph should open with the consequence: on these views, the question about LLMs is closed. Then the examples show why.
Comment 3 (line 23): "this is very vague and it makes it sound as though there are already people saying that the text is what's most important in philosophy. This is not true."
This is the deepest problem. My sentence presents the text-focused view as a pre-existing "conception" that the paper merely "takes up." But it's the paper's thesis — the thing Section 1 argues for. Nobody in the literature says "the text is what matters in philosophy" in quite this way. The paper CONSTRUCTS this position from Dellsén, Bengson, Williamson, Lipton, etc. So the introduction should say "We argue that..." — presenting it as the paper's own claim, not as received wisdom.
In Growing the Image: "We will argue that neither approach is satisfactory, and we propose thinking of Midjourney... as a new type of artistic medium." The paper's position is stated as the paper's position. In Agents of Change: "I argue here that a particular feature of agentive experience is in fact responsible." Same pattern.
Comment 4 (line 23): "the connection between this sentence and the one which precedes it is completely unclear."
The reader goes from "a philosophical contribution is assessed by examining the reasoning on the page" to "Floridi et al. argue that LLMs do not reason abductively." Why? What connects these? The connection is: IF evaluation is about the reasoning, THEN LLMs need to be able to reason well to do philosophy. Floridi says they cannot. But this logic is entirely implicit. The reader is left to reconstruct it. The fix: state the bridging question. "Whether LLMs can produce texts of the right kind is a further question." NOW the reader knows why Floridi and Zahavy are being introduced — they're addressing this further question.
Comment 5: Same problem with Zahavy.
Comment 6 (line 25): "these standards" — unclear referent. The previous paragraph never clearly stated what the standards are. The paper's thesis says evaluation bears on "the reasoning itself" but that's vague. "These standards" could refer to anything.
The fix is structural. If the paper's thesis is stated clearly, and the objections are clearly introduced as threats to it, then the thesis paragraph itself anchors everything.
Now let me look at what Nick's published introductions actually do in terms of structure.
Growing the Image: Block quote → here are two ways to characterise Midjourney (agent, tool) → we argue neither works → we propose medium. It's: framing → options → thesis.
Agents of Change: My aim → what's new → the rival view → my position. It's: aim → novelty → existing approaches → thesis.
Both of them: set up the problem, present existing options, state the paper's thesis.
This introduction should do the same: the LLM question → existing conceptions of philosophy (practitioner vs. text-internal) → the paper's thesis → the objections it faces.
The practitioner paragraph sets aside one conception. Then the paper states its own view: evaluation is text-internal. Then it introduces the threat: even if that's right, maybe LLMs can't reason well enough. Then it states its response: we argue they can.
That sequence works. Let me draft it.
For the transition sentence, I'm now thinking just: "Whether the same should be expected of philosophy depends on what you take philosophy to be." One sentence. Sharp. The next paragraphs unpack it.
For the practitioner paragraph, using Nick's suggested opening but in my own formulation: "On some conceptions, philosophy requires being a certain kind of subject, and the question of whether LLMs can do it is closed before it opens." Then Hadot, Wittgenstein, Merleau-Ponty, Nietzsche/Sorgner.
For the thesis paragraph: "We argue in this paper that philosophical evaluation bears on properties of the reasoning itself — on what is there in the text — and that whether LLMs can do philosophy is therefore a question about the texts they produce." This is the paper's own claim, stated as such. Then the bridge: "Whether LLMs can produce texts of the right kind is a further question." Then Floridi, Zahavy, consequence, thesis.
"We argue otherwise." — Short, after the buildup. Earns its place.
Then roadmap as its own paragraph.
Here is the new version from the change point. Everything before "One reason to be less optimistic" stays untouched.
---
Whether the same should be expected of philosophy depends on what you take philosophy to be.
On some conceptions, philosophy requires being a certain kind of subject, and the question of whether LLMs can do it is closed before it opens. For Hadot (1995), philosophy is a practice of self-transformation; for the later Wittgenstein (1953), a form of therapy; for Merleau-Ponty, it requires us to 'slacken the intentional threads which attach us to the world' (1945, p. xv) in order to examine them. On Nietzsche's account, as Sorgner reads it, philosophers are creators of values whose work expresses drives and a psychophysiology bound to human embodiment.[^3][^ac]
We argue in this paper that philosophical evaluation bears on properties of the reasoning itself — on what is there in the text — and that whether LLMs can do philosophy is therefore a question about the texts they produce.[^2] Whether LLMs can produce texts of the right kind is a further question. Floridi et al. (2024) argue that LLMs do not reason abductively: they produce plausible continuations rather than considered explanations. Zahavy (2026) argues that theoretical innovation requires embodied simulation of a kind that LLMs, operating entirely in symbols, cannot perform. An LLM that cannot reason its way to a good explanation will not produce philosophical texts worth assessing. We argue otherwise.
Section 1 argues that the standards by which philosophy is evaluated bear on properties of the reasoning itself. Section 2 presents the objections from Floridi et al. and Zahavy. Section 3 responds. Section 4 considers what a demonstration would look like.
---
What changed and why, comment by comment:
Comment 1 (transition sentence): "Whether the same should be expected of philosophy is less straightforward, because..." → "Whether the same should be expected of philosophy depends on what you take philosophy to be." One clause. The "because" construction was trying to give a reason and a consequence in parallel, which didn't parse. This just makes the claim.
Comment 2 (practitioner opening): "the philosophical work happens in the practitioner rather than in anything assessable in what she produces" → "philosophy requires being a certain kind of subject, and the question of whether LLMs can do it is closed before it opens." Front-loads the consequence for LLMs. The examples then show why the question is closed. Mirrors how Growing the Image states consequences immediately: "if a human didn't make the image, who did?"
Comment 3 (text-focused as existing consensus): "On other conceptions — the ones this paper takes up — a philosophical contribution is assessed by examining the reasoning on the page" → "We argue in this paper that philosophical evaluation bears on properties of the reasoning itself." Now stated as the paper's thesis, not as a pre-existing "conception." Mirrors "We will argue that neither approach is satisfactory" (Growing the Image) and "I argue here that" (Agents of Change). The paper's position is the paper's position.
Comments 4 and 5 (missing connective tissue): Floridi and Zahavy now follow a bridge sentence: "Whether LLMs can produce texts of the right kind is a further question." The reader can see why they are being introduced — they are reasons to doubt LLMs can produce the right texts. Before, they appeared from nowhere after a sentence about text-internal evaluation.
Comment 6 ("these standards" unclear): The closing paragraph no longer says "these standards." It says "We argue otherwise" — which refers clearly to the claim that LLMs cannot produce philosophical texts worth assessing.
The "neither paper concerns philosophy directly, but both, if sound, bear on it" meta-commentary is gone. The relevance is now shown by the structure: we said evaluation is about reasoning → can LLMs reason? → Floridi and Zahavy say no → we disagree.
Shall I update the note?
*Chi scrive un'introduzione deve dire dove sta andando — non dove stanno tutti gli altri.*
### User
Yeah update the note
### Assistant
Done.
*Ogni frase dell'introduzione dovrebbe farsi capire da sola — senza che il lettore debba indovinare i nessi.*
### User
I've made some more edits and there's still one big inline comment
### Assistant
## Step 1: Flagged text
> "We argue in this paper that philosophical evaluation bears on properties of the reasoning itself"
The comment: "this is an unfortunate and meaningless turn of phrase. Yeah, you need to do much better in this sentence about making the idea clearer to the reader because it's fucking dreadful right now."
## Step 2: Classify
Stating the paper's thesis.
## Step 3: Diagnose
**Casual/vague phrasing** — "bears on properties of the reasoning itself" is abstract to the point of emptiness. What properties? What reasoning? "Bears on" is doing no work — it means something like "concerns" or "is about" but vaguer. And "properties of the reasoning itself" sounds like it should be precise but is not. The reader gets: evaluation is about... some properties... of some reasoning... itself. That is not a thesis. A thesis should tell you what the paper claims, in terms specific enough that you could disagree with it.
The problem: the sentence is trying to compress the entire argument of Section 1 (philosophy is evaluated by the quality of the argument on the page — its explanatory power, how it handles objections, what understanding it affords) into one abstract phrase. But the compression empties it of content.
## Step 4: Published paragraphs
### Assistant
Three published thesis statements:
> We will argue that neither approach is satisfactory, and we propose thinking of Midjourney, which will be our case study, and similar text to image systems as a new type of artistic medium.
> — *Growing the Image*
The thesis names the specific thing being proposed ("a new type of artistic medium"). The reader knows what the paper claims.
> My aim here is to show how experience leads to the belief that time passes. The novelty of the account I propose lies in the type of experience which causes this belief.
> — *Agents of Change*
Again specific: "experience leads to the belief that time passes." Then: the novelty is the TYPE of experience. The reader knows what to expect.
> My aim in this paper is to show that we also hear empty spaces. We hear the volume of empty space enclosed by the walls of the church.
> — *Hearing Spaces*
The thesis is concrete: "we hear empty spaces." Then an example immediately ("the volume of empty space enclosed by the walls of the church"). The reader can picture what is being claimed.
What these share: the thesis tells the reader something SPECIFIC. "We propose a new type of artistic medium." "Experience leads to the belief that time passes." "We hear empty spaces." Each of these is concrete enough that you could disagree with it. "Philosophical evaluation bears on properties of the reasoning itself" is not — it sounds like it should mean something but does not tell the reader what the paper actually claims.
## Step 5: Rewrite
The thesis needs to say, specifically, what the paper argues. The paper argues: a piece of philosophy is assessed by examining what is on the page — how the case is built, whether it illuminates its subject matter, how it handles pressure. Who or what produced it is a separate question. If this is right, then whether LLMs can do philosophy depends on whether their texts exhibit these qualities.
That is a concrete, disagreeable claim. Let me compress it into a thesis sentence:
> We argue in this paper that a piece of philosophy is assessed by examining what is on the page — how the argument is built and whether it illuminates its subject matter — and that who or what produced it is a separate question.
That tells the reader the specific claim. "How the argument is built and whether it illuminates its subject matter" is concrete enough to disagree with (someone could say: no, philosophy is assessed by the biography of the philosopher, or by the tradition it belongs to). And "who or what produced it is a separate question" sets up the LLM application directly.
The full sentence in context:
> We argue in this paper that a piece of philosophy is assessed by examining what is on the page — how the argument is built and whether it illuminates its subject matter — and that who or what produced it is a separate question. Whether LLMs can do philosophy is therefore a question about the texts they produce.[^2]
## Step 6: Diff
**Before:**
> We argue in this paper that philosophical evaluation bears on properties of the reasoning itself — on what is there in the text — and that whether LLMs can do philosophy is therefore a question about the texts they produce.[^2]
**After:**
> We argue in this paper that a piece of philosophy is assessed by examining what is on the page — how the argument is built and whether it illuminates its subject matter — and that who or what produced it is a separate question. Whether LLMs can do philosophy is therefore a question about the texts they produce.[^2]
**Diagnosis:** Vague phrasing — "bears on properties of the reasoning itself" is abstract to the point of emptiness. **Fix:** replaced with specific content: "a piece of philosophy is assessed by examining what is on the page — how the argument is built and whether it illuminates its subject matter — and that who or what produced it is a separate question." The reader now knows what the paper claims.
Shall I apply this?
*La tesi di un articolo deve essere abbastanza chiara da poterla negare — altrimenti non dice nulla.*
### User
"We argue in this paper that a piece of philosophy is assessed by examining what is on the page — how the argument is built and whether it illuminates its subject matter — and that who or what produced it is a separate question. Whether LLMs can do philosophy is therefore a question about the texts they produce.2" this makes it sound as though the question of whether LLMs can do philosophy is secondary to our claim about philosophy being assessed on the page. You need to think harder about this. What's going on here is simple oh fucking hell... It's just so depressing you have to tell you the same things again and again. It's obviously what you've given me here is obviously wrong. Obviously.
### Assistant
You are right. The paper's thesis is not "philosophy is assessed on the page" with the LLM question as a corollary. The paper's thesis IS about LLMs. The claim about text-internal evaluation is the premise that makes the LLM question tractable — it is in service of the thesis, not the thesis itself.
The thesis is: LLMs can do philosophy. The argument for it goes: philosophy is evaluated by what is on the page → so provenance does not determine quality → so LLMs can produce philosophy if their texts are good enough → and they can.
The introduction should state the thesis (LLMs can do philosophy) and say what the argument turns on (the claim about text-internal evaluation), not the other way around. Currently it foregrounds the premise and backgrounds the thesis.
Something closer to: "We argue that LLMs can produce good philosophy. The argument turns on a claim about how philosophy is evaluated: if what we assess is the reasoning on the page — how the argument is built and whether it illuminates its subject matter — then who or what produced it is a separate question, and the texts LLMs produce are assessable by the same standards as anyone else's."
That puts the thesis first (LLMs can do philosophy) and the text-internal claim in its proper role (the ground on which the thesis stands). The reader knows from the start what the paper is arguing FOR, not just what it is arguing ABOUT.
Do you want me to try other versions of this, or is this the right direction?
*Il cuore di un articolo non è mai la premessa — è la conclusione che la premessa rende possibile.*
### User
I see what's got wrong here. Originally there were two paragraphs, one about philosophy which required a human. That remains. But the second one was entirely different. It was about the majority of philosophy in the analytic tradition. It's all evaluated by texts. You need to bring all of that information back. That's where it's gone wrong. Okay, and then once you do that, then that sets the ground for our proper claim about LLMs in the last paragraph of the introduction. wtf man, what a load of fucking shit.
### Assistant
You are right. The introduction needs three moves, not two:
1. Practitioner conceptions — philosophy requires a human, question closed (this is there)
2. Text-internal conceptions — in analytic philosophy, evaluation concerns what is on the page (this got CUT when I removed Dellsén/Bengson/Williamson to avoid redundancy with Section 1 — but the mistake was cutting the IDEA, not just the specific sources)
3. The paper's thesis — given (2), LLMs can do philosophy if their texts are good enough
What went wrong: I conflated "don't pre-empt Section 1 by naming every source and criterion" with "don't have a paragraph about text-internal evaluation at all." The introduction needs to ESTABLISH that analytic philosophy is evaluated on the page — that is the ground the paper stands on. It just should not do Section 1's work of developing the specific criteria (Lipton, Williamson, Bengson, etc.).
So the fix is: restore a paragraph between the practitioner paragraph and the thesis paragraph. This paragraph says: in analytic philosophy, contributions are evaluated by examining the texts themselves — how the arguments are built, whether they illuminate their subject matter, how they handle objections. It can mention that this is the mainstream position in recent analytic metaphilosophy without listing every source and every criterion. Then the thesis paragraph says: given this, LLMs can do philosophy if their texts meet these standards. We argue they can. Floridi and Zahavy give reasons to doubt it. We respond.
Do you want me to draft this restored three-paragraph structure, or do you want to talk through what the middle paragraph should say first?
*Tre gradini, non due — e il secondo è quello che regge il peso di tutto l'argomento.*
### User
I'm gonna have to talk you through it, aren't I? You can't be trusted otherwise.
### Assistant
Go ahead. I am listening.
*Chi ascolta prima di scrivere sbaglia meno di chi scrive prima di ascoltare.*
### User
No no, well I mean first of all you just need to find one of the previous version history drafts of this section and bring that paragraph back in. Verbatim and then we can just edit and move things around until it fits. That's obviously what should be done. You shouldn't have removed in the first fucking place. Waste of my fucking time.
### Assistant
Version 3 (15:09) is the one just before I cut the paragraph — it's the largest recent version. Let me read it.
### Assistant
That version doesn't have the original text-focused paragraph either — it already has my rewrite. Let me go further back to the pre-edit version.
### Assistant
There it is. The original paragraph:
> Some conceptions locate philosophy in texts. Dellsén et al. (2024) argue that philosophical progress consists in putting people in a position to increase their understanding — what they call the *for-whom rather than by-whom* account — in which public utility of the work, not the internal states of whoever produced it, determines whether progress has occurred. Bengson, Cuneo, and Shafer-Landau (2022) characterise philosophical inquiry as *theory construction* evaluated by criteria: accommodation of data, explanatory power, integration, theoretical virtue, all of which are assessable by examining the theory itself. Williamson (2024) defends an abductive methodology judging theories by their simplicity, elegance, and explanatory power. On any of these accounts, a philosophical contribution is a text exhibiting certain properties; the question who or what produced it does not enter the evaluation.[^2]
Let me restore it now — inserting it between the practitioner paragraph and the current thesis/Floridi paragraph, with the inline comments stripped since we will edit from here.
### Assistant
Paragraph restored (with "Bengson et al." fixed per the style rule). The introduction now has the three-paragraph structure: practitioner conceptions → text conceptions → thesis/Floridi. The thesis paragraph below it still needs fixing, but the ground is back.
*Il paragrafo tolto era il pavimento — senza di esso, tutto il resto galleggiava nel vuoto.*
### User
Okay, well, the paragraph you've just dropped in there, something like that content needs to be in there somewhere, but it's yeah, it needs a lot of work still. I've added a sentence to the very beginning of it which I think is closer to how I think it should begin. Maybe you have ideas on how to refine that though. Yeah, also you need at least another sentence after the one I've just put in there to make clear to the reader why you're talking about what you're talking about. As I say in the online comment, the Delson stuff is not described in a way that a reader is going to understand if they haven't read the original text. Basically you're right to say that it's putting people in a position to increase their understanding. And maybe you can use the for whom rather than by whom account, although you need a page number reference in line for that. But you don't bring it back to text. Okay, similarly when you move to Benggson you're not talking about the text. So basically you just haven't the at the moment the paragraph is not really explaining what this other half of ways of doing philosophy is and why we can sort of say well text is the the currency for these things. I mean at some point a little bit of levity can be put in early on as well, maybe as a footnote, maybe as part of the body text, about academics in philosophy being told to publish or perish. Okay, we're publishing texts, we're publishing words. So yeah, that paragraph is not great and needs severely rethinking about from first principles. Notice that bullet point I've put in immediately afterwards though. Again, that's just going to make it clearer to the reader what's going on here, right? We're saying, look, LLMs but if you if you have lots of an output-based text-based philosophy, then LLMs are not excluded in the way that they are by, like they are in Wittgensteinian style philosophy or whoever, Melopontian philosophy. But, and then this is a nice way of leading up to our final paragraph and therefore leading into the main text, this doesn't mean that they are capable of producing texts of sufficient quality. Okay, so that's how you do it. I mean I thought this was obvious but for some reason you're fucking flapping around again.
### Assistant
Good, I can see the structure now. Let me go back to the Dellsén source to get the characterisation right before drafting.
### Assistant
Let me find the page number for Dellsén's "for-whom rather than by-whom" phrase.
### Assistant
The "for-whom" phrase is at line 873, which falls between page headers 679 (line 856) and 680 (line 907). So it is on p. 679.
Now let me draft the revised paragraph and bridge. Here is what I am proposing — I will present it in the chat, not edit the note yet:
---
The text-based paragraph should read something like:
> More common in 21st Century analytic philosophy is what we might think of as an *output* based approach. Analytic philosophers publish arguments, and it is the published arguments that are assessed. Dellsén et al. (2024) make this explicit: philosophical progress, on their account, consists in putting people in a position to increase their understanding, and this happens not through private insight but by way of philosophical ideas — theories, arguments, distinctions — becoming publicly available (p. 679). The question is not who had the insight but whether the published work enables understanding. Bengson et al. (2022) and Williamson (2024) give this a methodological footing: philosophical theories are assessed by specific criteria — accommodation of data, explanatory power, integration, theoretical virtue — all of which bear on the text itself.[^2]
Then the bridge (replacing the bullet point):
> On these approaches, LLMs are not excluded from the start in the way they are on practitioner conceptions. Whether they can produce texts of the right kind is a further question.
Then the Floridi paragraph follows as it is.
The changes: the paragraph now starts from your opening sentence, adds a sentence saying what "output based" means in practice (philosophers publish arguments; the arguments get assessed), develops Dellsén using the actual source (publicly available ideas, not private insight — connecting to TEXT), compresses Bengson and Williamson into one sentence as methodological support rather than parading them separately, and closes with the assessment bearing on the text itself. The bridge sentence then makes the logical move: on these conceptions, LLMs are in the game; the question is whether they can play.
Shall I apply this?
*La moneta della filosofia analitica è il testo pubblicato — e chi lo ha coniato non compare sulla faccia.*
### User
Okay, your new version is much better and I do want you to add it to the note. I don't like it when you say Delsin et al. make this explicit. Because it sounds like they're arguing for what I'm arguing for and they're not. So that needs to be changed. Would you be able to add a tiny bit on at the end just to make what you're saying about well the Banks and Williamson stuff I think should be there, but could is there I feel like there's one more piece of content that could be added around this part of a paragraph. Any thoughts? Moving on to your bridge, why don't you just use my version as far as it goes and then build on that? Why why change the words?
### Assistant
You are right on all three points.
On Dellsén: "make this explicit" implies they are arguing for the same thesis as the paper. They are not — they are arguing about what philosophical progress IS, and the paper draws on their account. Something like "Dellsén et al. (2024) argue that philosophical progress consists in putting people in a position to increase their understanding" — just state what they argue, no framing about what it "makes explicit."
On the missing piece: I think what is absent is the consequence of all three sources taken together — namely that on these accounts, philosophical quality is something a reader can assess from the text without knowing anything about its author. That is the thought that connects the output-based paragraph to the LLM question. The closing sentence gestures at it ("the question who or what produced it does not enter the evaluation") but it arrives too quickly. A sentence before it drawing the threads together — something like: what these accounts share is that the criteria they describe are satisfied, or not, by what is on the page — would give the closing sentence something to land on.
On the bridge: you are right, I should use your words. Your bullet says "On these more text based approaches, LLMs are not excluded automatically" — that is clearer than my version and it is yours. I will use it verbatim and build from there.
Let me apply all of this now.
### Assistant
Done. Your bridge sentence kept verbatim, Dellsén stated without "make this explicit," the shared consequence drawn out before the closing sentence, and Bengson/Williamson compressed into one sentence rather than paraded.
*Tre fonti, un punto in comune — e il punto in comune è ciò che regge il paragrafo.*
### User
Okay, good. I just remembered in the same part of that paragraph about the "They're satisfied by what's on the page, not by who produced it." There should be some mention in that sentence or in the immediate succeeding sentence about blind review. Okay, and then yeah, and I also thought maybe we could put as a footnote, didn't we have something in one of the previous versions which mentions some academic scandal where some big cheese managed to get something published because of their name rather than what's on the page? Look back in previous versions of this and you should see what I mean because that should go in there somewhere. Maybe it's in the ce v ce ce v version of this entire text.
### Assistant
Let me search for that academic scandal reference in the conversation history and previous versions.
### Assistant
Not in those versions. Let me search the conversation file from earlier today where the CEV analysis was done — that is where it would be.
### Assistant
I cannot find the academic scandal / big cheese reference in any of the conversation files or previous versions. It may have come up in an earlier session than the ones from today. Do you remember any more about it — was it a specific philosopher, or a specific journal? The Sokal hoax (a physicist getting nonsense published in a cultural studies journal) is the most famous case, but that is science-into-humanities, not philosophy-into-philosophy. There is also the case of journals accepting papers from prestigious authors without proper review — but I do not want to fabricate a specific case.
In the meantime, let me do the blind review edit. The current sentence reads: "What these accounts share is that the criteria they describe are satisfied, or not, by what is on the page; the question who or what produced it does not enter the evaluation." Blind review should be woven into that sentence or the one immediately after. Something like:
> What these accounts share is that the criteria they describe are satisfied, or not, by what is on the page. Blind review — the norm across analytic philosophy — gives this thought an institutional form: referees are asked to assess what a paper does, not who produced it.[^2]
That adds blind review as the institutional expression of the text-internal principle. The footnote about the academic scandal can be added to the blind review sentence once we find or decide on the right example.
Shall I apply the blind review addition now, and we can sort out the footnote separately?
*La revisione cieca esiste perché la disciplina ha già deciso che il testo viene prima del nome.*
### User
Sokal was indeed the scandal I was thinking of. It's definitely been in in some sort of version or some sort of note. Maybe just do a search for this word and you should be able to find it either in the chats or in the vault somewhere else okay and you should see where it is and secondly your thing about blind review I don't write in I don't write sentences like that that's not in my style at all and yeah Yeah. Not quite
### Assistant
Found it. The Sokal point from the plan: "The Sokal hoax worked in a domain where evaluation was impressionistic. It probably couldn't work in a top analytic philosophy journal, because the referees would check the arguments. This isn't because analytic philosophers are smarter — it's because the evaluative norms of the discipline operate at the artefact level."
That is the point you want as a footnote — it connects blind review to the text-internal claim. The Sokal hoax succeeded where evaluation was impressionistic; it would not succeed in analytic philosophy because the arguments are checked. The evaluative norms operate on the text.
Now, on my blind review sentence — you are right, "gives this thought an institutional form" is not how you write. Let me think about this differently. The blind review point and the Sokal point could both sit in a footnote rather than the body text, since the body text already says "the question who or what produced it does not enter the evaluation." Blind review is the institutional evidence for that claim. A footnote:
> [^X]: Blind review is the institutional expression of this: referees assess what a paper does, not who produced it. The Sokal hoax is instructive here — it succeeded in a domain where evaluation was impressionistic. It would be unlikely to succeed in a leading analytic philosophy journal, because referees check the arguments. The evaluative norms of the discipline operate on the text.
That keeps the body text clean and puts the evidence in a footnote where it supports the claim without interrupting the paragraph's flow.
Or — if you want blind review in the body text, it could be woven into the closing sentence more naturally: "...satisfied, or not, by what is on the page, which is why the discipline evaluates them under blind review; the question who or what produced it does not enter the evaluation."
Which direction do you prefer — footnote with both blind review and Sokal, or blind review in the body with Sokal in a footnote?
*Il caso Sokal dimostra il contrario di ciò che sembra — non che la revisione sia debole, ma che dove è forte, il testo basta.*
### User
"That is the point you want as a footnote — it connects blind review to the text-internal claim. The Sokal hoax succeeded where evaluation was impressionistic; it would not succeed in analytic philosophy because the arguments are checked. The evaluative norms operate on the text." yeah, that's the idea I was thinking of, although I don't think it's very well articulated in those in that little bit of text. Think about the scandal and then try and develop the CEV of this footnote.
### Assistant
The Sokal affair shows something specific about where text-internal evaluation holds and where it does not. Sokal got a nonsensical paper into *Social Text* because the evaluative norms there were impressionistic — the paper sounded right, used the right vocabulary, cited the right people, and that was enough. Nobody checked whether the physics made sense, because the evaluation was not operating on the arguments. The hoax worked precisely because there was a gap between how the text sounded and what it actually said.
The point for this paper: in analytic philosophy, that gap is much smaller, because referees check the arguments. Blind review is not just a fairness measure — it reflects the fact that the discipline's standards bear on the reasoning itself. When a referee reads a paper under blind review, they are assessing whether the distinctions do real work, whether the objections are handled, whether the position earns its reach. If those things are in the text, the paper is good; if they are not, no reputation will save it. The Sokal hoax could not easily succeed in that environment, because the evaluative norms are argument-checkable. Where they are argument-checkable, the surface IS the depth — what you see on the page is what there is to assess.
As a footnote, I think it should read:
> Sokal's deliberately nonsensical paper was accepted by *Social Text* in 1996 because the journal's evaluative norms were impressionistic — the paper deployed the right vocabulary and cited the right authorities, and that was enough. The hoax would be difficult to replicate at a leading analytic philosophy journal, because referees check the arguments. Blind review in analytic philosophy is not merely a fairness measure; it reflects the fact that the discipline's evaluative standards bear on the reasoning itself, and the reasoning is there on the page for anyone to assess.
Three sentences, no framing, direct entry. Do you want this attached to the "the question who or what produced it does not enter the evaluation" sentence, or somewhere else in the paragraph?
*Il caso Sokal funzionò dove il giudizio era di superficie — e fallì dove qualcuno avrebbe letto l'argomento.*
### User
Blind review should be mentioned in the body, so-call should be a footnote. Your so-call footnote was fucking awful because you presumed that the reader's gonna know what you're talking about, so you didn't explain what the actual scandal was. It's also a very you've written it in a very very judgmental and value-laden terms, which is A not in keeping with analytic philosophy, B just shit and not how I would write anyway, as you should know by now.Please start the task(s) again completely from scratch, keeping this in mind.
### Assistant
You are right. Let me start from scratch on both the blind review sentence and the Sokal footnote.
The body text currently reads: "What these accounts share is that the criteria they describe are satisfied, or not, by what is on the page; the question who or what produced it does not enter the evaluation."
Blind review needs to go into the body near this sentence. The Sokal footnote needs to explain what the affair actually was — a reader who has not heard of it should understand the point — and it needs to do so without value-laden language.
Let me read a published paragraph to ground the voice, then draft both.
### Assistant
For the body text, blind review should be woven into the existing sentence naturally. The point: analytic philosophy already evaluates texts without knowing who wrote them — that is what blind review is.
Body text revision:
> What these accounts share is that the criteria they describe are satisfied, or not, by what is on the page. Analytic philosophy already evaluates work this way: under blind review, referees assess what a paper does without knowing who produced it.[^br]
Then the Sokal footnote, attached to this sentence. The footnote needs to: (1) explain what the Sokal affair was — a physicist submitted a paper, it was a hoax, and the journal published it; (2) say why it worked — the evaluation did not operate on the arguments; (3) draw the contrast with analytic philosophy — where referees check the arguments, the hoax could not work in the same way.
Sokal footnote:
> [^br]: In 1996, the physicist Alan Sokal submitted a paper to *Social Text*, a cultural studies journal, in which he argued that quantum gravity is a social and linguistic construct. The paper was a hoax — Sokal had written it to test whether a journal would publish an article that, as he later put it, "sounded good" but whose arguments were nonsensical (Sokal 1996a, 1996b). The journal published it. The affair is instructive because it turned on a gap between how a text sounds and whether its arguments hold up. In a discipline where referees check the reasoning — where they test whether distinctions do real work and whether objections are handled — the gap is smaller. Blind review in analytic philosophy operates at this level: the referee reads the argument, not the author's name, and the argument is what is assessed.
That explains the affair to a reader who has never heard of it, without value-laden language — no "nonsensical," no "impressionistic," just a description of what happened and why it is relevant. The contrast with analytic philosophy is drawn through what referees do, not through a judgement about which discipline is better.
Shall I apply both to the note?
*Chi spiega uno scandalo senza giudicare lascia al lettore il compito di trarre le conclusioni.*
### User
"it turned on a gap between how a text sounds and whether its arguments hold up. In a discipline where referees check the reasoning — where they test whether distinctions do real work and whether objections are handled — the gap is smaller. Blind review in analytic philosophy operates at this level: the referee reads the argument, not the author's name, and the argument is what is assessed." I like the footnote up until this bit. This bit is all dreadful. Better is something like the following. Such a case. It seems that in such a case, the work was not judged by the quality of a text on the page. Also was the SoCal thing even anonymous? Because if not, that needs to be emphasized much more as well.
### Assistant
You are right — the Sokal submission was not anonymous. Sokal submitted under his own name as a professor of physics at New York University. The journal knew exactly who he was. That sharpens the point considerably: the paper was accepted with the author's credentials visible, and still nobody checked the arguments.
Here is the revised footnote:
> In 1996, the physicist Alan Sokal submitted a paper to *Social Text*, a cultural studies journal, in which he argued that quantum gravity is a social and linguistic construct. The paper was a hoax — Sokal had written it to test whether a journal would publish an article that, as he later put it, 'sounded good' but whose arguments were nonsensical (Sokal 1996a, 1996b). The journal published it. The submission was not anonymous: Sokal's name and institutional affiliation were attached. In this case, the work was not assessed by the quality of the reasoning on the page.
The last sentence does what you suggested — states the point and lets the reader connect it to the body text themselves. The non-anonymous detail is emphasised because it makes the contrast with blind review sharper.
Shall I apply both the body text (blind review sentence) and this footnote to the note?
*Sokal firmò con il proprio nome — e nessuno lesse l'argomento.*
### User
" The journal published it. The submission was not anonymous: Sokal's name and institutional affiliation were attached. In this case, the work was not assessed by the quality of the reasoning on the page." you can really do better than this. First of all, the first sentence is short and stubby in a way I would never do. Seconds don't describe don't say the submu well say Basically what needs to be emphasized in here is that it would seem that the journal was focusing on the person rather than what's on the page. Okay, so the idea of this footnote is we're, you know, maybe not directly, but at least implicitly alluding to the subject-based versions of philosophy which we mentioned in that previous paragraph. Now come the fuck on man.
### Assistant
You are right. The footnote should connect to the practitioner paragraph — the journal evaluated the person, not the text, and that is precisely the kind of evaluation the practitioner conceptions would endorse. Here is the revised ending:
> In 1996, the physicist Alan Sokal submitted a paper to *Social Text*, a cultural studies journal, in which he argued that quantum gravity is a social and linguistic construct. The paper was a hoax — Sokal had written it to test whether a journal would publish an article that, as he later put it, 'sounded good' but whose arguments were nonsensical (Sokal 1996a, 1996b). The journal published it under Sokal's own name, with his institutional affiliation attached — suggesting that what was being assessed was not so much the reasoning on the page as the person behind it.
One sentence does the work of three. "Not so much the reasoning on the page as the person behind it" echoes the distinction between practitioner and output conceptions without naming it. The reader makes the connection.
Shall I apply?
*La firma contava più dell'argomento — e questo dice tutto su come veniva giudicato il testo.*
### User
Okay, that seems much better to me. Can you now double double double check online that this is an accurate characterization of what happened in this particular hoax? If it is, you need and to add it as a footnote in the repo of this section.
### Assistant
Good. The Wikipedia article confirms every claim in my footnote. The relevant facts:
1. 1996 — confirmed
2. Submitted to *Social Text*, "an academic journal of cultural studies" — confirmed
3. Paper proposed that "quantum gravity is a social and linguistic construct" — confirmed
4. Sokal was "a physics professor at New York University" — confirmed, his name and affiliation were on the paper
5. Not anonymous — confirmed. The article explicitly says "based on Sokal's credentials *Social Text* published the article"
6. The journal "did not practice academic peer review at the time" and "did not submit the article for outside expert review by a physicist" — confirmed. This is an important detail: *Social Text* did not peer-review at all. So it was not just that they failed to check — they did not have a peer review process.
7. Sokal's stated purpose: to test whether the journal would "publish an article liberally salted with nonsense if (a) it sounded good and (b) it flattered the editors' ideological preconceptions" — confirmed
One correction needed in my footnote: the paper was published partly "based on Sokal's credentials" — the Wikipedia article makes this explicit. That strengthens the point about the person being assessed rather than the text.
The footnote is accurate. Let me now apply both the blind review body text and the Sokal footnote.
### Assistant
Done. Blind review in the body text, Sokal footnote with the full explanation of the affair, the non-anonymous submission emphasised, and the closing that connects it to the practitioner/text distinction.
I also added the detail from the Wikipedia article that *Social Text* "did not practise peer review at the time" — which sharpens the point further.
*Chi non legge l'argomento giudica la firma — e chi giudica la firma non giudica affatto.*
### User
Okay, I'm moving back onto section one now. You'll notice I've removed some of the shittier sentences in the first paragraph. None of this attribution or who named it first, nonsense or anything like that. That's all gone. I've I don't like the phrase in paragraph two. Second sentence, he was not trying to get his readers to see familiar, he was trying to get his readers to see familiar philosophical materials differently. Don't know what that means, it's very vague and very stupid. So that needs to be changed the next clause. Yeah the next clause is very ugly and very unclear as well. You need to slow down and make this much much more clear to the reader because right now it's fucking bollocks.
### Assistant
The flagged text is: "Ryle was not reporting a previously unknown item in the world, he was trying to get his readers to see familiar philosophical materials differently, and what he achieved consists in the argument itself — the distinctions it draws, and what it opens for others to assess and contest."
Two problems. "Trying to get his readers to see familiar philosophical materials differently" — what does that mean? It is vague. What was Ryle actually doing in 'Ordinary Language'? He was arguing that certain philosophical puzzles arise from misunderstanding how ordinary expressions work. That is specific. "See familiar philosophical materials differently" is a description of any philosopher who has ever argued for anything.
Then "what he achieved consists in the argument itself — the distinctions it draws, and what it opens for others to assess and contest" — this is rushed and unclear. "Consists in the argument itself" is the claim, but then the dash opens into a parenthetical that says two things at once (distinctions + opening for others) without developing either.
The paragraph needs to say, specifically, what Ryle was doing and why it is different from Watson and Crick. Watson and Crick found an arrangement that was already there. Ryle was not finding anything that was already there — he was making a case. The case IS the contribution. If the case had been different, the contribution would have been different, in a way that does not apply to Watson and Crick (where a different paper describing the same arrangement would have been the same discovery).
Let me draft a replacement for the flagged portion:
> Ryle was not reporting a previously unknown item in the world; he was making a case — arguing that philosophical puzzles about the nature of ordinary language arise from confusions that can be dissolved by attending to how expressions actually function. Had he made a different case, drawn different distinctions, the contribution would have been different. The work consists in the argument as made on the page.
That is specific about what Ryle was doing (not vague "seeing materials differently"), and it makes the contrast with Watson/Crick explicit (different case = different contribution, unlike different paper = same discovery).
Then the Pigliucci qualification follows naturally: "This is not to say that philosophy starts from nothing..."
Shall I apply this?
*Se non sai dire cosa Ryle stava facendo, non puoi dire perché è diverso da Watson e Crick.*