Observation
Which task, input, version, and response were actually recorded?
AI ETHICS
Linguistic self-reports and convincing behavior count as evidence, but alone they do not establish subjective experience.

What can language, behavior, self-reports, and technical architecture warrantedly support about possible subjective experience—and what statement remains responsible while a certain diagnosis is unavailable?
A small learning software provider tests a text-based practice assistant. In a demonstration, the system explains a pun and changes its answer after additional context is entered. The team can describe the performance, but does not know whether subjective experience follows from it. The case is entirely fictional and does not describe a particular product.
Which task, input, version, and response were actually recorded?
Which ability or inner state is inferred from it? Which counterhypothesis explains the same data?
Which theory of consciousness makes a feature relevant, and what technical evidence is still missing?

Phenomenal experience must be distinguished from cognitive access. Intelligence, emotion, self-description, and responsibility are likewise not interchangeable terms. A lack of evidence establishes neither consciousness nor its absence.
A clarifies terms, aim, and assumptions. B opens up competing explanations. D marks uncertainty, missing data, and limits. C formulates a narrower, action-guiding question. The sequence A → B → D → C replaces a hasty yes/no diagnosis.
Exploration identifies role and knowledge gaps. Reflection and Analysis set evidence alongside the opposing position. Decision and Recommendation establish a provisional rule. Feedback and Evaluation state when it will be reassessed.
R is used only for a specific specialist question as an expert handoff, for example to review technical documentation or a measurement design. It structures the request and feedback. It does not replace K and does not measure consciousness.
The research sources in the article support different philosophical or empirical subquestions. Human studies are not presented as AI tests. F, K, and R organize questions and decisions; their descriptions do not establish methodological effectiveness. This overview is not a consciousness diagnosis, a validated audit standard, or legal advice.