History · AI · Truth & Epistemics October 4, 2026 10 min read

How to Use AI for Historical Research Without Inventing the Past

A magnifying glass examines layered historical papers beneath a faint digital network.

Ask AI for the exact speech Pope Urban II delivered at Clermont in 1095, and the question already assumes there is one recoverable text waiting to be retrieved. A polished answer may satisfy that assumption before examining it.

Historical research often begins by refining the question. Here, we need to ask which account we are reading, when it was written, and what it can tell us about the event. Those questions matter whether the first answer comes from a chatbot, a textbook, or a website.

Use AI to organize questions and candidate claims, then check those claims against identifiable passages. Keep the writer, document, translation, and uncertainty visible in the final prose. The aim is a useful account whose strength matches its evidence.

We will work through a small set of claims about Clermont, then turn the results into a revised paragraph. For broader context, see our introduction to Urban II at Clermont and First Crusade narrative.

Begin with a question the sources can answer

“What exactly did Urban say?” and “How do surviving writers describe his appeal?” are different research questions. The first asks for recoverable wording. The second asks us to compare reports.

Georg Strack's study explains that the speech is known through differing chronicle accounts rather than Urban's own surviving text. He examines their forms and compares them with traditions of papal oratory. This is historical interpretation, not authentication of every sentence in a preferred account. Strack, 2012, pp. 30–31

A manageable starting question is therefore: What can we responsibly say about the appeal at Clermont, and which details need attribution or qualification?

Write that question down before asking for a summary. It gives you a boundary for deciding which sources to inspect and which claims belong in the finished paragraph. You can broaden the inquiry later without pretending that a small source packet settles the whole history of the crusades.

Build a source packet before building a narrative

The Internet Medieval Sourcebook collection offers accessible English extracts. Its six numbered entries comprise five narrative extracts and a papal letter. They are not six interchangeable transcripts.

For this exercise, the collection provides the starting texts. Selected scholarship supplies context and interpretation; publication records help check dates and editions. A publisher's description can provide bibliographic information, but citing it does not mean the book's full argument has been examined.

Record four dates separately when available: the event, the document's composition, the translation or edition, and your access. A website's recent update does not make its translation recent or its chronicler contemporary with the event.

For each item, also record who wrote it, its genre, the translator, and an inspectable location. “Section 2, Robert the Monk, crowd-response passage” is more useful than a bare collection link. A page number is useful only when you identify the edition it belongs to.

Then ask what the document was trying to do. Recounting events, exhorting an audience, giving instructions, and interpreting earlier texts are different activities. Treating them all as neutral recordings erases information you need to judge them.

Audit the sentences before trusting the paragraph

Consider this deliberately weak statement:

Six independent eyewitness transcripts preserve Urban's exact speech, confirming every detail of the familiar account.

This is a constructed teaching example, not a recorded answer from a named AI model. No model benchmark or error rate is being reported here.

The sentence bundles several claims: a count, independence, witness status, transcript status, exact wording, and agreement. One citation cannot be assumed to support all six. Split the sentence and ask what evidence each part requires.

The table shows what changes after checking the passages. Its first column contains constructed overclaims, not quotations from the sources.

Constructed overclaims and their source checks
Illustrative overclaimWhat the inspected evidence supportsBetter wording
The surviving account gives Urban's exact words.Strack discusses differing reports, not a recoverable verbatim text. Discussion, p. 30Specify which writer reports the detail.
Robert wrote his account as the speech happened.The modern edition's publisher gives probable completion around 1110. Edition descriptionRobert wrote later; qualify precise dating.
The Flanders letter has a certain December 1095 date.Cowdrey calls it undated, conventionally assigned to late December 1095. Reprint, p. 27Preserve the conventional dating as a qualification.
Fulcher's silence proves Jerusalem was absent from Urban's plan.The supplied Fulcher account does not name Jerusalem; that alone cannot establish intent. Section 1Report the omission without turning it into proof.
No council records survive because no transcript survives.Cowdrey discusses surviving canons in a version preserved by Lambert of Arras. Reprint, p. 28Distinguish council records from a speech transcript.

Some rows require rejection; others require narrower wording. That difference matters. A claim can be too strong without being wholly invented. Your audit should identify the actual problem rather than attach one blanket label to everything.

Keep the writer between you and the event

“Urban said” makes a claim about the speaker. “Robert reports that Urban said” makes the chronicler's mediation visible. It is a small grammatical change with a substantial effect on what you are asserting.

In the translated collection, Robert reports a collective cry after the address. Baudri's account describes varied reactions. Their presentations should remain attributable rather than being combined into a scene that appears unanimously documented. Robert and Baudri extracts, sections 2 and 4

Quotation adds another layer. If you copy English words exactly, you are quoting a translation of a writer's report. In this collection, Robert's extract is credited to Dana C. Munro's 1895 translation, pages 5–8. A quotation should retain that attribution. It cannot establish Urban's exact original-language wording. Translation credit, section 2

A vivid allegation also needs careful handling. A document's authenticity does not independently verify every event it alleges. When researching accusations or atrocity rhetoric, separate what the text claims from what other evidence establishes. This exercise does not independently investigate those allegations.

Compare sources without counting copies as confirmations

Two websites can reproduce one translation. Two historical writers can depend on an earlier narrative. Agreement can therefore arise through transmission as well as independent observation.

Strack discusses later writers' relationships with the Gesta Francorum. He also distinguishes its broader account of Urban's preaching in France from a word-for-word rendering of the Clermont speech. A source's relationship to the event must be examined before it is counted as another witness to a specific utterance. Strack, pp. 31, 34–37

Draw a simple map: which website reproduces which edition, and which texts are reported to draw on others? Mark relationships you have checked and leave uncertain ones uncertain. You do not need to pretend you have reconstructed the complete manuscript tradition to notice that three links may lead back to one source.

Comparison also means preserving different emphasis. An omission, an incompatible statement, and an inaccessible passage are three different findings. Write each one accurately. “Not found in the material I checked” is a statement about your search; it is not a universal statement about the past.

Rewrite only as far as the evidence allows

After the audit, a cautious paragraph might read:

Urban II's appeal at Clermont in 1095 survives through differing reports rather than a recoverable verbatim transcript. Robert's account was written after the event, while the Flanders letter offers a different kind of evidence about the undertaking. Details should be attributed to their sources, and differences between accounts retained instead of blended into a supposedly exact speech.

This paragraph is our source-informed synthesis, not a translated medieval passage. Its distinctions draw on Strack's discussion, the Robert edition description, and Cowdrey's analysis of the letter.

The revision removes the claim to exact wording, identifies the different kinds of evidence, and retains attribution. Further historical argument remains possible. Useful qualification locates uncertainty precisely, without spreading vague doubt over every sentence.

That is one practical application of the site's Truth Engine Framework: make the path from assertion to support inspectable.

Give AI a task you can check

Once you have opened the sources, AI can be assigned bounded organizational work: propose a list of claims, compare supplied passages, suggest narrower wording, or identify questions for further investigation. Check its output at each step, including the passage locations it proposes.

Here is a reusable prompt. Replace the brackets with your material, and give every source a stable label of your own.

Research question: [one precise question]
Source packet: [opened texts, authors, dates, editions,
translators, and passage locations]
Proposed paragraph: [text to inspect]

Using only this packet:
1. Split the paragraph into individual claims.
2. Match each claim to a source and inspectable passage.
3. Separate reported speech, translation quotations,
   paraphrase, and historical interpretation.
4. Flag differences in dates, genre, and source dependence.
5. Mark missing support as NOT FOUND IN THIS PACKET.
   Do not invent quotations, citations, dates, or witnesses.
6. Suggest narrower wording and list unresolved questions.

I will independently check every proposed source mapping.

The last sentence describes the researcher's responsibility. It does not make the model's mappings correct. A tidy table can contain a mistaken source attribution just as a fluent paragraph can contain a mistaken fact.

Keep a working ledger with five fields: claim, source, passage location, limitation, and revised wording. For example: “Robert wrote in 1095” → modern-edition publisher description → description paragraph → approximate dating only → “Robert's account was written later, probably around 1110 according to the publisher.” Leave unsupported claims marked for further research.

This is why the distinction between a prompt and a framework matters in practice. The prompt proposes work. The surrounding process determines what you accept and how you check it.

Before you use the result

The Library of Congress's primary-source guide encourages observation, reflection, questions, further investigation, and revision through comparison. Those habits provide a useful educational foundation for the workflow here; the AI-specific ledger is our adaptation. Teacher's Guide: Analyzing Primary Sources

The American Historical Association's 2025 guidance offers a related educational example: a starter bibliography requires checking every reference and looking beyond the generated list. Its policy table is adaptable classroom guidance, not a rule governing every history publication. AHA guidance, scope statement and Appendix 2

Before using your own paragraph, check:

  • Have I opened the cited sources and found the relevant passages?
  • Have I distinguished the event from the account and its translation?
  • Have I considered whether apparent confirmations share a source?
  • Have I retained disagreements and stated the limits of my search?
  • Has the final edit preserved the qualifications in my ledger?

Start with one paragraph from your next historical query. Trace its most consequential claim all the way to the passage that supports it. Keep the result and the unresolved questions together.

Frequently asked questions

Can I use AI to suggest historical sources?

Yes, as a starting list to investigate. Check that each work exists, find the relevant passage, and look beyond the suggestions. A correct title can still be attached to a claim it does not support.

Does a primary source prove what happened?

It provides evidence that must be interpreted in context. The author, purpose, date, genre, and relationship to other material affect what you can responsibly infer. A reported statement and an independently established event are different claims.

Will this prompt eliminate hallucinations?

No such guarantee is established here. The prompt helps organize a review whose outputs still require checking. This article presents a bounded documentary exercise, without measuring any model's reliability.

Research note: This article uses translated extracts, selected scholarly passages, and publication records. It does not claim manuscript collation, consultation of the complete modern critical editions, or an exhaustive literature review. The weak claims are constructed illustrations; no separate AI-model experiment was run.

AI disclosure: This article was drafted with AI assistance and reviewed against the cited material. Readers should check the referenced passages when reusing consequential claims.

Continue the conversation

Get the next dispatch.

Occasional essays, tool notes and fiction updates. Confirm by email and unsubscribe at any time.