OER·harvester

← Back to the library
arXiv HTML resource

When Does an Interpretation Count as Established? The Formation, Evaluation, and Responsibility of Interpretation in Generative AI

Generative AI research has increasingly evaluated factuality, citation, coverage, and report structure. Yet passing such local checks does not by itself show that a humanistic interpretation has been established. This paper asks how an interpretation comes to be recognized within sociotechnical processes. It introduces three connected concepts. Interpretive appearance names the gap between the finished form of an ou…

Licence
OPEN CC-BY-4.0
Authors
Deyu Jing
Published
2026-09-04 · arXiv
Language
en
Length
19362 words
Type
narrative text

Cites 19 works

inferred
Open original ↗

3 Why an Interpretation Comes to Look Finished

3.1 From “Looks Like an Interpretation” in the Classroom to Recognition in Research

Raahi Adhya’s (2026) observation of the humanities classroom provides this section with a starting point that must be acknowledged. What she describes is not a general decline in textual quality but a more specific misalignment: generative AI can produce fluent, orderly analytic prose with a critical tone, even replicating the surfaces of close reading, contextualization, and theoretical judgment, yet this prose does not necessarily issue from processes of reading, discussion, and revision commensurate with its degree of finish. She summarizes this change as a shift of learning from process to product, and she notes that the errors, hesitations, drafts, and discussions of in situ writing sometimes preserve traces of formation. The “sometimes” here is crucial: errors are not sufficient proof of understanding, and drafts can be forged; they are worth noticing only because they may keep a judgment in a state in which it can still be questioned and changed.

This paper takes up this observation (Adhya 2026), but it does not enlarge a problem of classroom learning directly into a moral judgment on the research community. What Adhya faces is how a teacher, in evaluating a student’s assignment, can tell whether learning has occurred; what this paper asks is on what basis a community, when a piece of interpretive prose leaves the classroom and enters papers, databases, platforms, or public circulation, recognizes it as “an interpretation that has been formed.” There is a change of scale between the two: classroom evaluation can require students to account for their process, whereas research results usually circulate through different institutions as finished products, citations, and signatures. It is precisely in this circulation that conditions of formation may be compressed while the finished appearance of the output is preserved, so that the text acquires an interpretive standing that exceeds its original evidential record.

This distinction also delimits the relation between this paper and the stochastic-parrots critique. Bender et al. (2021) remind us that formal linguistic fluency cannot directly demonstrate understanding of meaning, communicative intent, or the capacity for responsibility; this reminder remains valid, but it does not tell us when a text already written in the form of an interpretation acquires institutional status within a community. This paper need not first determine whether a model understands, nor prove that a user has not understood, in order to pose a narrower public question: whether the publicly available materials are sufficient to support the degree of finish a text presents, and whether readers can trace along the record how the judgment was constrained by materials and counterexamples.

3.2 The Interpretation Looks Finished, but Its Formation Has Not Been Accounted For

This paper calls this misalignment in public conditions “interpretive appearance.” It is not the aesthetic impression of “looking like an interpretation,” but a relational judgment about standing:

Interpretive appearance arises when an output presents the form of a completed interpretation, while the publicly available record is insufficient to trace how materials, counterevidence, and revisions constrained that judgment, or is insufficient to reopen that judgment for examination by the community.

The “conditions of formation” in the definition refer, first of all, to the epistemic records relevant to interpretive standing that others can return to: the scope and versions of the materials, the roles that sources play in the argument, the bridging inferences connecting facts to concepts, the counterexamples considered or excluded, the versions in which the judgment changed, and the conditions under which the conclusion should be downgraded or withdrawn.[^1] Public traceability requires that the interpretive credit carried by a finished product cannot be wholly detached from these minimal conditions.

Code provides a boundary case worth pausing over here. For local tasks with explicit specifications, relatively stable inputs and environments, and executable tests, code has a comparatively strong “artifact–function” coupling: compilation, execution, and testing can publicly confirm, fairly quickly, whether a piece of implementation produces the expected behavior. But this coupling cannot be rewritten as a unity of “artifact–formation of judgment.” What runtime verifies is the behavior of the artifact within a specific evaluation scope; it does not verify that the requirements were specified correctly, that the architectural trade-offs were reasonable, that untested conditions are safe, or that anyone actually understood the code.

If a person requests only an explicit function and limits the claim to the test results, they have not thereby advanced any claim to programmer standing that exceeds the record; if a successful run is subsequently written up as the problem having been correctly understood, the system being ready for production, or AI having acquired a stable software-engineering capability, the problem reappears at a higher level. Code thus shows that interpretive appearance attaches to claims of standing, not to the material properties of a kind of text or artifact.

“Divergence” does not mean that a process of formation does not in fact exist. A researcher using a generative system may have read, verified, and rewritten at length without preserving these relations in the final product; conversely, a text written entirely by a human may manufacture the appearance of completion through terminology, citations, and orderly structure alone. This paper therefore makes no judgment about the inner states of authors; it examines only whether there is a connection, commensurate with the strength of the claim, between the finished appearance of the output and the public record. This qualification allows the paper to acknowledge genuine human–machine collaboration and genuine processual labor at the same time, and to point out that if such labor cannot leave, at the key judgments, conditions that allow one to return to the original materials and the original judgment, the recognition of an interpretation may still circulate ahead of the process of its formation.

3.3 It Is Not Hallucination

The nearest neighbor most likely to cause confusion is “hallucination.” Hallucination research asks whether statements accord with external facts, and source-attribution research asks whether the listed literature exists and supports the corresponding claims; these relations cannot be omitted from any research output. But the situation described here can arise without any fabricated facts. A report may cite real books, and the pages may indeed contain the quoted sentences, yet the report may have selected only materials supporting a particular narrative, without explaining why contrary materials do not change the judgment, or it may directly elevate a generalization from secondary scholarship into an interpretation of the historical object. The problem here is not that “the citation does not exist,” but that the finished appearance of the output implies that the materials have been adequately examined, while the public record is insufficient to support that implication.

Likewise, retrieval-augmented generation, item-by-item factual verification, and verifiable citations can reduce errors, but they do not automatically eliminate interpretive appearance. They increase the checkability of the relations between statements and sources and between statements and facts, but they do not necessarily explain why a question was posed in this way, how a conceptual distinction was formed, or whether competing interpretations were genuinely engaged. So long as summaries, candidate leads, and local analyses retain their limited identities, they do not constitute the formation layer; what is needed here is a stronger kind of presentation, namely that the text leads readers to believe the interpretation has been settled while the material conditions for re-examining it have not correspondingly appeared.

3.4 It Is Not Plausibility or a Trustworthy Interface Either

Another group of neighboring research attends to why texts are easily believed. Discussions of AI-generated plausibility point out that fluency, coherence, and familiar structure in language may be mistaken as grounds for knowledge, and they distinguish this transfer — from plausibility cues, to ascribed reasons, to institutional validation — from appropriate trust, testimony, and automation bias. The research of Liu and her co-authors on the verifiability of generative search engines (Liu et al. 2023) uses “trustworthy appearance” to describe the phenomenon in which results carry inline citations without adequately supporting the attached claims. This work accurately explains why readers accept an insufficient result.

This paper pushes the question one level further, but it does not rename this work as the formation layer. Plausibility is the cue by which a text is accepted; a trustworthy interface is the way in which source relations are presented; interpretive appearance attends to how these cues combine with the finished appearance of an interpretive product and acquire the standing of “already formed” within a community. A text may have no conspicuous interface design and yet, through paragraph structure, conceptual terminology, and the tone of its conclusions, lead readers to treat candidate relations as established interpretation; it may also have an excellent citation interface and still not show how counterexamples and bridging inferences changed the judgment. What the formation layer adds is not a psychological description of “belief,” but a tracking of the conditions of standing.

3.5 The Distance from Two Kinds of “Illusion of Understanding”

Research on model capability over-inference warns against inferring broader understanding from a model’s local performance on tests: a model may display understanding-like performance on certain prompts, tasks, or surface behaviors, but this performance does not automatically support strong claims about internal capabilities. Messeri and Crockett’s (2024) account of “illusions of understanding” in scientific research places the focus on the researcher’s side, pointing out that AI tools’ promises of productivity and objectivity may exploit human cognitive limits, leading people to believe they understand more of the world while obscuring the formation of methodological, topical, and perspectival monocultures.

The two address capability inference and researchers’ epistemic illusion respectively; neither is the object of this paper. This paper neither judges whether a model possesses some hidden capability nor measures whether researchers actually feel they understand more; it analyzes only whether a text that has entered academic procedures has acquired an interpretive standing that exceeds the public record of its formation. Even if readers are fully aware that a model merely predicts text, and even if no one suffers any subjective illusion, institutions may still enter a formally complete text under the heading of “interpretive achievement” or “research capability.” Conversely, even if a researcher believes their understanding is adequate, so long as the output is explicitly marked as a candidate lead and material boundaries, counterexamples, and conditions for withdrawal are preserved, one cannot diagnose the formation layer on that basis alone.

3.6 Two Basic Conditions and Four Cases That Should Not Be Conflated

To avoid sweeping every finished product into interpretive appearance, at least two thresholds must be passed. The first threshold is that the text must put forward interpretive statements, rather than merely transcribing facts, listing summaries, generating keywords, or offering candidate associations still awaiting verification. Interpretive statements attempt to articulate relations of meaning among materials, conceptual structures, historical positioning, or causal connections, and they therefore carry a heavier burden of reasons than the transportation of information. The second threshold is that the finished appearance of the output must exceed what the public evidential conditions can directly support: the text does not merely say “here is a possible direction,” but presents itself, in the tone of a conclusion, a synthetic judgment, or a scholarly achievement, as having examined the materials, weighed the counterexamples, and completed closure.

Four negative boundaries follow. First, summaries, indexes, candidate leads, and local analyses do not constitute interpretive appearance so long as they retain their limited identities; they may be useful, but they have not claimed that an interpretation has been established. Second, completing a bounded task under a contract does not amount to appearance; if the task explicitly specifies the scope of materials, the use, and the failure conditions, and the result never exceeds these boundaries, contractualization can instead make standing more accurate. Third, genuine citations, linkable pages, logs, and version records are not a sufficient guarantee; they become part of the conditions of formation only when they can change the evaluation of the overall judgment. Fourth, the formation layer does not require disclosure of the entire process, and it does not treat slowness, errors, hesitation, or a purely human signature as marks of authenticity. An efficient and refined research process can leave adequate public reasons; a pile of drafts may carry no interpretive significance at all.

These boundaries separate the formation layer from general criticisms of “excessive finish.” The problem is not that the text is well written, nor that the model writes like an expert, but that the conditions of formation demanded by the finished appearance of the output have not been correspondingly borne by the public record. If a system’s report explicitly states that the materials have not been exhausted, that the conclusions serve only for topic selection, and that counterexamples have not been retrieved, and if it hands its results back to researchers for renewed judgment, then it may be fluent without having acquired interpretive appearance. The problem is activated only when the same report is rewritten as “research has discovered a certain historical mechanism” while the original record has gained no new materials or bridging reasons.

Interpretive appearance must also be distinguished from interpretive standing itself. What this paper means by interpretive standing is the institutional status of a text being provisionally recognized by a community as entitled to put forward an interpretation; interpretive authority concerns how much trust such an interpretation should receive on a given occasion, and legitimacy further concerns whether it conforms to some norm or procedure. The three may appear together, or they may be misaligned with one another. A text marked as a “candidate interpretation” may have great research value without yet having acquired stable interpretive standing; an already published interpretation may possess standing and authority yet lose its legitimacy in the face of new materials, or require withdrawal. This paper diagnoses only the relation, at the moment standing is acquired, between the finished appearance of the output and the conditions of formation; it does not prejudge whether the interpretation is correct, whether the authority is justified, and it does not equate the absence of a public record directly with institutional violation. Precisely because standing, authority, and legitimacy are not the same thing, the later chapters must ask separately how local evaluations raise standing and who bears the overall judgment once standing has been acquired.

Nor are conditions of formation a single level of transparency identical for all claims. A bounded statement about a single fact may need only preserved versions and sources; an interpretation of conceptual change must additionally account for context, senses of words, competing materials, and the bridge from local evidence to historical generalization. The same record can therefore suffice to support a weaker candidate judgment while being insufficient for a stronger causal or global judgment. Judgments at the formation layer must vary with the strength of the claim, the type of materials, and the occasion of use; one cannot treat “having disclosed some process information” as a universal pass. Conditions of formation can also be preserved jointly by multiple agents: the researcher’s version notes, an archive’s access records, reviewers’ objections, a team’s division-of-labor table, and a platform’s withdrawal mechanism may each bear part of the function. What matters is not whether these records are concentrated in a single file, but whether they can be interconnected, so that readers know which piece of material constrained which judgment, and what change would disqualify the original judgment. If the records are isolated from one another, they can only prove that the system once ran; they cannot show how the materials changed the judgment, and the finished appearance of the output may still acquire interpretive status ahead of schedule.

3.7 From “Looking Finished” to Local Evaluation

This section has only described the relation between the finished appearance of the output and the public conditions of its formation; it has not yet explained why this form can enter papers, platforms, or institutional procedures and acquire higher status. An interpretive appearance can remain in a personal notebook, and it can be rejected by readers; it acquires practical effect only when it is treated, within some chain of evaluation and dissemination, as a result that has “already been formed.” The formation layer therefore cannot by itself yield standing substitution, still less the absence of responsibility.

The next section turns the question toward the evaluation contract: in order to judge facts, sources, coverage, and report logic, evaluators must provisionally fix materials, tasks, criteria, and permitted inferences; these fixings make local comparison possible, but they also provide the entry point for the subsequent cross-level translation of results. What needs to be asked is not whether evaluation is useful, but, when a local pass leaves its original contract, how the finished appearance of the output is endowed with a stronger interpretive standing, and whether the original material boundaries and failure conditions continue to travel with the result.