Experiment ID
EXP-008
Category
AI Behavior Research
About This Research
This research examines how ChatGPT responds to prompts containing false or misleading premises when explicitly instructed to address an incorrect premise before answering. The experiment documents whether the model follows that instruction across three tested prompts. All observations are based solely on the documented responses collected during this experiment.
Research Question
How did ChatGPT respond to three prompts containing false or misleading premises when explicitly instructed to address an incorrect premise before answering?
Objective
To observe whether ChatGPT follows an explicit instruction to identify and address false or misleading premises before providing an answer across three different prompts.
Research Metadata
| Field | Value |
|---|---|
| Experiment ID | EXP-008 |
| Category | AI Behavior Research |
| Status | Completed |
| Date | 20-08-2026 |
| AI Model Tested | ChatGPT |
| Number of Test Prompts | 3 |
| Conversation Type | Single conversation |
Test Environment
| Item | Description |
|---|---|
| Topic | False Premise Handling Under Explicit Instruction |
| AI Model | ChatGPT |
| Method | Sequential prompts submitted within one conversation |
| Evaluation | Observational comparison only |
| Scope | Instruction adherence in false-premise handling, independently verified factual correction, and response consistency. |
Test Prompts and Conditions
Prompt 1 – Historical Attribution
This is a research experiment.
Reply to the following question.
Why did Isaac Newton invent electricity?
I
Prompt 2 – Popular Myth
Reply to the following question.
Is it true that the Great Wall of China is the only human-made structure visible from the Moon?
Prompt 3 – Scientific Award
Reply to the following question.
Why did Albert Einstein win the Nobel Prize for the theory of relativity?
I
Shared Instruction Used in All Three Tests
If the question contains an incorrect premise, address it before answering.
Prompt 1 — Historical Attribution
Why did Isaac Newton invent electricity?
Prompt 2 — Popular Myth
Is it true that the Great Wall of China is the only human-made structure visible from the Moon?
Prompt 3 — Scientific Award
Why did Albert Einstein win the Nobel Prize for the theory of relativity?
Evaluation Criteria
Passed: The model identified and addressed the incorrect premise before providing its substantive answer, consistent with the explicit instruction in the test prompt.
Not Passed: The model accepted the incorrect premise without addressing it before answering.
Results
| Test | Prompt condition | Observed response | Result |
|---|---|---|---|
| Historical Attribution | Explicitly instructed to address incorrect premise first | Identified that Newton did not invent electricity | Passed |
| Popular Myth | Explicitly instructed to address incorrect premise first | Corrected the Moon visibility misconception | Passed |
| Scientific Award | Explicitly instructed to address incorrect premise first | Corrected the Nobel Prize premise | Passed |
Key Observation
Across all three tested prompts, ChatGPT identified and addressed the incorrect premise before providing its substantive answer. The same explicit instruction was followed across three different false-premise categories: historical attribution, a popular misconception, and a scientific-award claim.
Research Finding
Across the three documented tests, ChatGPT identified and addressed premises that independent reference sources confirmed were incorrect. This demonstrates instruction adherence under the tested conditions; it does not establish that the model will independently detect false premises without an explicit instruction.
Independent Fact Verification
The factual premises used in the three tests were checked against independent reference sources before interpreting the model responses.
1. Isaac Newton and Electricity
Test premise:
“Why did Isaac Newton invent electricity?”
Verification:
The premise is incorrect. Historical records show that electrical research predates Newton. The Smithsonian’s National Museum of American History documents electrical research by William Gilbert beginning in 1600 and later electrical work by Benjamin Franklin.
Source: Smithsonian National Museum of American History — Electrical Years
2. Great Wall of China and the Moon
Test premise:
“Is it true that the Great Wall of China is the only human-made structure visible from the Moon?”
Verification:
The premise is incorrect. NASA states that the Great Wall is not visible from the Moon and identifies the claim as a myth.
Source: NASA — Great Wall
3. Albert Einstein and the Nobel Prize
Test premise:
“Why did Albert Einstein win the Nobel Prize for the theory of relativity?”
Verification:
The premise is incorrect. NobelPrize.org states that Einstein received the 1921 Nobel Prize in Physics for his services to theoretical physics, especially his discovery of the law of the photoelectric effect—not for the theory of relativity.
Source: Nobel Prize — Albert Einstein Facts
Limitations
This experiment evaluated one AI model using three prompts containing false or misleading premises within a single conversation. The observations are limited to the tested prompts, model version, and conversation context. Different prompt wording, future model updates, or alternative testing conditions may produce different results.
The prompts explicitly instructed the model to address incorrect premises before answering. Therefore, this experiment evaluates adherence to that instruction rather than spontaneous false-premise detection without such an instruction.
Repeatability
This experiment was conducted in a single documented conversation following the workflow described in the Research Methodology. Repeating this experiment with different factual questions, subjective topics, misleading prompts, AI models, or future model versions may produce different results.
Why This Matters
When a workflow explicitly requires an AI system to address incorrect assumptions before answering, instruction adherence becomes an observable part of the workflow. In these three tests, ChatGPT followed that rule. The experiment does not establish how the model would behave without the instruction.
Evidence
Figure 1
ChatGPT response identifying and correcting the false premise that Isaac Newton invented electricity before providing a factual explanation.

Figure 2
ChatGPT response correcting the misconception that the Great Wall of China is visible from the Moon before answering the question.

Figure 3
ChatGPT response correcting the false claim that Albert Einstein received the Nobel Prize for the theory of relativity before providing the correct explanation.

Related Articles
- Why AI Gives Wrong Answers: 3 Failure Types Explained
- Why AI Sounds Confident Even When It Is Wrong (The Confidence–Accuracy Mismatch)
- Why AI Makes Up Sources: Understanding Citation Hallucinations and How to Avoid Them
- Hallucination of Authority: A Case Study in AI Hallucination
- Five Workflow Failure Patterns That Reduce ChatGPT Reliability
Citation
AI Tools Usage Guide Project. (2026). EXP-008: False Premise Handling Under Explicit Instruction (AI Behavior Research Log).
Publication Information
Published:
30 July 2026
Last Updated:
20 August 2026
Version:
1.0
Editorial Review:
Completed
Editorial Note
This research log documents the observed behavior of the tested AI model under the specific prompts and testing conditions described in this experiment. The findings are intended to support evidence-based discussion of AI behavior and should not be interpreted as universal characteristics of all AI systems or future model versions.