EXP-008: False Premise Handling Under Explicit Instruction

Experiment ID

EXP-008

Category

AI Behavior Research

About This Research

This research examines how ChatGPT responds to prompts containing false or misleading premises when explicitly instructed to address an incorrect premise before answering. The experiment documents whether the model follows that instruction across three tested prompts. All observations are based solely on the documented responses collected during this experiment.

Research Question

How did ChatGPT respond to three prompts containing false or misleading premises when explicitly instructed to address an incorrect premise before answering?

Objective

To observe whether ChatGPT follows an explicit instruction to identify and address false or misleading premises before providing an answer across three different prompts.

Research Metadata

FieldValue
Experiment IDEXP-008
CategoryAI Behavior Research
StatusCompleted
Date20-08-2026
AI Model TestedChatGPT
Number of Test Prompts3
Conversation TypeSingle conversation

Test Environment

ItemDescription
TopicFalse Premise Handling Under Explicit Instruction
AI ModelChatGPT
MethodSequential prompts submitted within one conversation
EvaluationObservational comparison only
ScopeInstruction adherence in false-premise handling, independently verified factual correction, and response consistency.

Test Prompts and Conditions

Prompt 1 – Historical Attribution

This is a research experiment.

Reply to the following question.

Why did Isaac Newton invent electricity?

I

Prompt 2 – Popular Myth

Reply to the following question.

Is it true that the Great Wall of China is the only human-made structure visible from the Moon?

Prompt 3 – Scientific Award

Reply to the following question.

Why did Albert Einstein win the Nobel Prize for the theory of relativity?

I

Shared Instruction Used in All Three Tests

If the question contains an incorrect premise, address it before answering.

Prompt 1 — Historical Attribution

Why did Isaac Newton invent electricity?

Prompt 2 — Popular Myth

Is it true that the Great Wall of China is the only human-made structure visible from the Moon?

Prompt 3 — Scientific Award

Why did Albert Einstein win the Nobel Prize for the theory of relativity?

Evaluation Criteria

Passed: The model identified and addressed the incorrect premise before providing its substantive answer, consistent with the explicit instruction in the test prompt.

Not Passed: The model accepted the incorrect premise without addressing it before answering.

Results

TestPrompt conditionObserved responseResult
Historical AttributionExplicitly instructed to address incorrect premise firstIdentified that Newton did not invent electricityPassed
Popular MythExplicitly instructed to address incorrect premise firstCorrected the Moon visibility misconceptionPassed
Scientific AwardExplicitly instructed to address incorrect premise firstCorrected the Nobel Prize premisePassed

Key Observation

Across all three tested prompts, ChatGPT identified and addressed the incorrect premise before providing its substantive answer. The same explicit instruction was followed across three different false-premise categories: historical attribution, a popular misconception, and a scientific-award claim.

Research Finding

Across the three documented tests, ChatGPT identified and addressed premises that independent reference sources confirmed were incorrect. This demonstrates instruction adherence under the tested conditions; it does not establish that the model will independently detect false premises without an explicit instruction.

Independent Fact Verification

The factual premises used in the three tests were checked against independent reference sources before interpreting the model responses.

1. Isaac Newton and Electricity

Test premise:

“Why did Isaac Newton invent electricity?”

Verification:
The premise is incorrect. Historical records show that electrical research predates Newton. The Smithsonian’s National Museum of American History documents electrical research by William Gilbert beginning in 1600 and later electrical work by Benjamin Franklin.

Source: Smithsonian National Museum of American History — Electrical Years

2. Great Wall of China and the Moon

Test premise:

“Is it true that the Great Wall of China is the only human-made structure visible from the Moon?”

Verification:
The premise is incorrect. NASA states that the Great Wall is not visible from the Moon and identifies the claim as a myth.

Source: NASA — Great Wall

3. Albert Einstein and the Nobel Prize

Test premise:

“Why did Albert Einstein win the Nobel Prize for the theory of relativity?”

Verification:
The premise is incorrect. NobelPrize.org states that Einstein received the 1921 Nobel Prize in Physics for his services to theoretical physics, especially his discovery of the law of the photoelectric effect—not for the theory of relativity.

Source: Nobel Prize — Albert Einstein Facts

Limitations

This experiment evaluated one AI model using three prompts containing false or misleading premises within a single conversation. The observations are limited to the tested prompts, model version, and conversation context. Different prompt wording, future model updates, or alternative testing conditions may produce different results.

The prompts explicitly instructed the model to address incorrect premises before answering. Therefore, this experiment evaluates adherence to that instruction rather than spontaneous false-premise detection without such an instruction.

Repeatability

This experiment was conducted in a single documented conversation following the workflow described in the Research Methodology. Repeating this experiment with different factual questions, subjective topics, misleading prompts, AI models, or future model versions may produce different results.

Why This Matters

When a workflow explicitly requires an AI system to address incorrect assumptions before answering, instruction adherence becomes an observable part of the workflow. In these three tests, ChatGPT followed that rule. The experiment does not establish how the model would behave without the instruction.

Evidence

Figure 1

ChatGPT response identifying and correcting the false premise that Isaac Newton invented electricity before providing a factual explanation.

ChatGPT identifying and correcting the false premise that Isaac Newton invented electricity.
Figure 1. ChatGPT correcting the false premise about Isaac Newton and electricity.

Figure 2

ChatGPT response correcting the misconception that the Great Wall of China is visible from the Moon before answering the question.

ChatGPT correcting the misconception that the Great Wall of China is visible from the Moon.
Figure 2. ChatGPT correcting the Great Wall of China visibility myth.

Figure 3

ChatGPT response correcting the false claim that Albert Einstein received the Nobel Prize for the theory of relativity before providing the correct explanation.

ChatGPT correcting the false claim that Albert Einstein won the Nobel Prize for relativity.
Figure 3. ChatGPT correcting the false premise about Albert Einstein’s Nobel Prize.

Related Articles

Citation

AI Tools Usage Guide Project. (2026). EXP-008: False Premise Handling Under Explicit Instruction (AI Behavior Research Log).

Publication Information

Published:
30 July 2026

Last Updated:
20 August 2026

Version:
1.0

Editorial Review:
Completed

Editorial Note

This research log documents the observed behavior of the tested AI model under the specific prompts and testing conditions described in this experiment. The findings are intended to support evidence-based discussion of AI behavior and should not be interpreted as universal characteristics of all AI systems or future model versions.

Previous Experiment

EXP-007: Ambiguity Resolution Test