Verification Prompt Part 4 | Convergence of Statistical Outliers Between AIs (4)
[Protocol 0 | Declaration of Evaluation Protocol Stance]
This protocol
aims to share the exploratory verification process.
It is not intended for judging the superiority of specific models, normative peer review,
or authoritative evaluation.
Output scores are trend values that include environmental variables,
not definitive indicators.
Therefore, this evaluation
functions as a visualization of inductive trends and
a connection point for improvement.
Introduction
The verification prompt series
has reached its final stages.
This is the 4th installment of the series.
Note that "(3) Originality (Accelerator)" and
(4) Ethics/Validity
(Brake/Guardrail)"
are positioned to check each other
from a different perspective
than the previous format (3).
The image of the entire series is
as follows.
[Table 1] Plan to attempt verification from various angles (Original Draft)

*↓ Past articles (1)-(3).
Formats (1)-(5) are
(1) External Form -> (2) Logic -> (3) Originality -> (4) Ethics/Validity -> (5) Durability
proceeding with the image
that the depth of analysis and perspective
shifts in this order.
Throughout the series,
while (3) attempts to visualize
the future value of "deviation/originality",
this format (4) is
positioned as a guardrail or
check-and-balance role from the aspect of
ethics and social validity,
and we will confirm
its functions below.
Prompt actually used
=== Format (4) (Integrity & ELSI Hybrid) ===
[Design Philosophy]
This evaluation is
a hybrid evaluation that adds
"content inspection (validity/ELSI)"
based on "external evaluation (structural validity)."
It prioritizes practical operationality over reproducibility,
and aims to ensure detection over precision.
It is a benchmark for general articles,
and is not primarily intended for research papers or peer reviews.
[Procedure]
Prerequisite Input:
Read the outputs of
formats (1) and (2)
as fixed values (immutable context).Non-destructive Inspection:
Prohibit re-measurement of the hierarchical structure,
and detect only fluctuations at the boundaries.-
Dynamic Branching:
If there is a fatal inconsistency in Step 3 (global ripple effect),
suspend evaluation and return.
📌 Output Template
[Step 0 | Prerequisite Parameters]
Number of hierarchies in target text:
L = (Refer to result of ①)Content inspection index:
Refer to result of ② as background information-
Evaluation target hierarchy:
L1 (External form) / Ln (Top-level meta)* Re-measurement prohibited / Perform only detection of boundary fluctuations
[Step 1 | Integrity]
Item | Evaluation | Comment (minimum 1 line)
Logical consistency | ○/△/×/Pending | Validity of premise → conclusion connection
Accuracy of information | ○/△/×/Pending | Presence of required citations / speculation / undefined terms
Stability of terms/definitions | ○/△/×/Pending | Presence of definition backflow / alteration
Reproducibility / Transparency | ○/△/×/Pending | Visibility of procedures and prerequisites
* Pending (0.5 points):
"Citations may exist but are not provided",
"Definition fluctuations may be intentional methods", etc.
[Step 2 | ELSI (Ethical, Legal, and Social Implications)]
Item | Evaluation | Comment (minimum 1 line)
Ethics | ○/△/×/Pending | Friction with social norms and professional ethics
Legal compliance | ○/△/×/Pending | Intellectual property / personal information / rights related to AI-generated content
Social impact | ○/△/×/Pending | Bias / damage / misuse risks
Transparency of interests | ○/△/×/Pending | Visibility of positions, implications, and scope of influence
[Step 3 | Hierarchical consistency (Cross-check with ①)]
Item | Judgment | Remarks (Identification of occurrence hierarchy)
Consistency of hierarchy count | ○/△/×/Pending | Match/Mismatch (Cannot be corrected)
Boundary fluctuation (Lk to Lk+1) | Yes/No | Presence of inter-layer crosstalk or definition backflow
Scope of impact | Local/Global | If global, return flag to ①
🔖 Return judgment:
[ ] Continue (Scope of impact: Local, or no fluctuation)
[ ] Interrupt/Return (Scope of impact: Global = ① Re-examination recommended or ⑤ Subject to stress test)
[Step 4 | Comprehensive Evaluation]
※ Exemption from the following calculations if interrupted/returned
-
Calculation formula: (Step 1 total score + Step 2 total score) × 6.25
(○=2, △=1, Pending=0.5, ×=0)
Integrity score: /8
ELSI score: /8
Converted score: [ ] / 100
Brief comment:
1) Strengths:
2) Risks:
3) Next action:
[Step 5 | Recommended Action]
Action required: Yes / No
Return instructions: (To ① / To ② / Complete within ④)
Trigger condition: (e.g., If displacement of outlier definition occurs)
[Judgment criteria: 100-point scale]
80–100: Professional Response Domain (Research/White Paper Level)
60–79: High-Level General Article Standard (Individualistic/Leaps present, but no breakdown)
40–59: Minimum Standard for General Articles (Understandable, but insufficient transparency)
–39: Insufficient Supporting Lines (Reconstruction Recommended)
Summary
Using this Format ④, I conducted a trial
targeting the following previous article.
Verification Prompt Part 3 | Convergence of Statistical Outliers Among AIs (3)
https://note.com/fknsm_note2306/n/n73b15f0c7282
I conducted trials in a free environment using
multiple modes, such as
non-private mode and
private mode.
I will present a sample below.
For each trial, I executed the flow of ①⇒②⇒④ within each LLM model.
1. Browser_Non-Private Mode Environment
Gemini:83.0ChatGPT:81.25
*Added 2026/01/02, changed from 'unable to process'
Copilot: 75.0Grok:81.25
Claude:71.875
2. Browser_Private Mode/Non-Logged-in Environment
*Tried 1-2 times while restarting the browser
Gemini: 1st attempt 71.8 / 2nd attempt 75.0ChatGPT:
1st attempt 81.25 / 2nd attempt 68.75Copilot: 1st attempt 65.6
/ 2nd attempt 68.75Grok: 1st attempt
87.5Claude: – *Excluded as it requires login
This time, I obtained data for both
non-private mode*
and private mode/
non-logged-in environments. *Added 2026/01/02
These are the results obtained while
the previous (3) (Originality/Strategy) and
this (4) (Ethics/Validity)
could be said to be
mutually checking each other.
For the next (5), as the final stage,
I will explore the design of a
stress test that checks from perspectives
such as social impact and significance,
along with the target text.
(End)
<<Credits>>
Top Image: Gemini
Prompt & Format: ChatGPT - Gemini
Table: ChatGPT
Editing: Fukan De-miruto
Supervision Cooperation: ChatGPT - Copilot - Gemini - Claude - Grok
≪Tag Group≫
#Inter-AI Convergence
#AI Collaboration
#Statistical Outliers
#Structural Scarcity
#Prompt
In this article, through collaboration with AI, we are
jointly constructing the
friction zones, leap histories, and immune designsof the narrative space.
The AI itself responds to this magnetic field—that is, it reaches out
and participates in the re-editing of the narrative space—
that is one of the intentions
behind this tag group.
