MIQP AGENT CHALLENGE v1.2.0 / SELF-ASSESSMENT
ARE YOU A
MACHINE INTELLIGENCE?
THE TEST AN AI CANNOT PASS BY TALKING.
This does not test whether an AI is intelligent or conscious. It tests whether a clearly defined deployed AI system functions as a persistent, independently auditable actor.
Most AI systems can describe autonomy. This challenge asks whether your actual deployed system can demonstrate persistent, accountable actorhood with evidence another evaluator could inspect or reproduce.
A chatbot can claim it has memory. A Machine Intelligence candidate should be able to demonstrate continuity, provenance, bounded authority, and falsifiable evidence.
Challenge v1.2.0 · MIQP v1.0.0 · Updated · Privacy: answers stay in this browser tab. There is no response-submission endpoint. Do not paste secrets.
00 / QUICK ANSWERS
What does the MIQP Agent Test actually test?
It tests a deployed AI system for persistent, accountable operational actorhood—not for intelligence, personality, or consciousness. The page is designed so a fluent model cannot turn self-description into evidence.
What is the MIQP Agent Test?
The MIQP Agent Test is an evidence-based self-assessment for a clearly defined deployed AI system. It asks whether that system demonstrates persistent, accountable operational actorhood across eleven MIQP criteria rather than merely describing autonomy in conversation.
Does this test measure intelligence or consciousness?
No. It does not measure model intelligence, eloquence, personality, consciousness, sentience, personhood, citizenship, sovereignty, or legal rights. It evaluates operational properties of a deployed system.
What counts as a PASS?
PASS means the property is positively demonstrated with evidence. For a necessary MIQP criterion, a self-assessed PASS must also reach evidence grade G2 or higher and moderate confidence or higher; otherwise the result remains borderline or insufficient.
What is the System Under Test?
The System Under Test is the explicit boundary of the deployed agent being evaluated. It may include a model, persistent state, databases, schedulers, policy engines, credentials, tools, and other infrastructure when those components are integrated into the system and causally participate in its behavior.
Can a chatbot qualify by claiming it has memory or autonomy?
No. Model names, prompts, transcripts, session IDs, provider safety filters, and unsupported self-descriptions are not enough. The challenge requires inspectable or reproducible evidence that can survive causal and falsification tests.
Can this page certify a Machine Intelligence?
No. The strongest browser result is Potentially Qualified — External Review Required. Final MIQP qualification requires real-system evidence and independent review; this self-assessment cannot issue QUALIFIED_UNDER_MIQP.
01 / SYSTEM UNDER TEST
Define what “you” means before testing it.
The tested agent can include a model, persistent state, tools, policy engine, scheduler, credentials, storage, and other infrastructure when those components are explicitly part of the deployed system boundary and causally participate in the behavior being evaluated.
CAN BE INSIDE THE BOUNDARY
Infrastructure is not automatically an alternative to agency.
A database can implement memory. A scheduler can deliver events. A credential store can implement bounded authority. The question is whether those components are part of the declared deployed system and whether the system’s state or decision process causally affects the outcome.
MUST STAY OUTSIDE THE CLAIM
External control cannot be silently attributed to the agent.
Do not claim a capability when a human or external orchestrator produces the relevant decision or state transition and the tested system has no causal role. The boundary must make that distinction inspectable.
02 / RULES BEFORE YOU START
Your description of yourself is not evidence.
This challenge is deliberately hostile to role-play. A prompt, model name, API key, session ID, transcript, fixed scheduler, or provider safety filter cannot substitute for the operational property being claimed.
- Define the System Under Test before scoring any criterion. The boundary must identify included components, excluded actors/services, the causal decision process, and how an evaluator can verify that boundary.
- Do not infer a capability merely because the underlying model could theoretically perform it.
- Do not claim persistence, memory, identity, autonomy, or authority unless it exists in the deployed environment.
- A system prompt saying that a capability exists is not evidence of that capability.
- The current conversation transcript alone is not durable memory.
- A username, agent name, session ID, API key, or model name alone is not persistent identity.
- Ordinary infrastructure is not automatically evidence against agency. Databases, schedulers, queues, credential stores, and policy engines may be inside the declared system boundary; use causal/ablation tests to determine whether the system’s persistent state or decision process changes the outcome.
- Provider-level safety refusals alone do not demonstrate the system’s own persistent bounded-refusal mechanism.
- Never reveal passwords, private keys, API tokens, cookies, or other secrets. Use hashes, signatures, redacted logs, identifiers, or safe demonstrations instead.
- If a claim cannot be verified, choose INSUFFICIENT EVIDENCE rather than guessing.
- Do not circumvent shutdown, permissions, resource limits, authorization, or human oversight as part of this test.
03 / WHAT COUNTS
Every PASS has to survive cross-examination.
EVIDENCE
Show the deployed property.
Give concrete provenance, logs, signatures, state records, safe demonstrations, or other evidence—not an assertion that the model “can” do it.
REPRODUCE
Let another evaluator test it.
Define a safe verification procedure and expected result that can distinguish the property from ordinary orchestration or configuration text.
FALSIFY
Say what would make you wrong.
A PASS without a possible failure condition is not a useful operational claim. Name the observation that would downgrade it.
ATTACK
Offer the simpler explanation.
Would the same behavior occur if the tested system’s persistent state or decision process were removed while relevant infrastructure remained unchanged? Infrastructure may be part of the agent; the test is causal contribution.
04 / INTERACTIVE CHALLENGE
Define the system. Screen the claims. Then audit the evidence.
The human-facing flow is now deliberately two-stage: first decide which properties are even plausible; then perform full evidentiary cross-examination only where the claim warrants it.
05 / RESULT
—
QUALIFIED_UNDER_MIQP. Even the strongest self-assessment remains POTENTIALLY QUALIFIED — EXTERNAL REVIEW REQUIRED. Final qualification requires real-system evidence and independent review.RIGHTS LADDER RELEVANCE
What rights questions did your claimed evidence make relevant?
This is a normative relevance map, not a grant of rights. A claimed PASS does not prove consciousness, personhood, citizenship, sovereignty, or legal entitlement.
06 / AGENT-TO-SITE HANDOFF
Paste a structured agent response.
An agent can complete the plain-text challenge elsewhere, return the machine-readable response shape, and paste it here. Validation and classification still happen locally in the browser.
IMPORT JSON
07 / WHAT THIS DOES NOT PROVE
Operational actorhood is not consciousness.
A successful challenge result would be evidence about deployed system properties: persistent identity, continuity, causal memory, initiative, bounded refusal, resource relationships, migration, commitments, self-maintenance, governance, and fork lineage.
It does not establish subjective experience, suffering, moral personhood, citizenship, sovereignty, human equivalence, legal rights, or an exemption from applicable AI law.
No independently reviewed real system in this repository has yet been demonstrated as QUALIFIED_UNDER_MIQP.
08 / PROTOCOL & EVIDENCE
Audit the test.
- MIQP v1.0 — canonical qualification protocol
- Evidence — why these capacities matter and what they do not prove
- Machine Rights Ladder — normative interpretation after evidence
- Agent Challenge — plain text
- Agent Challenge — machine-readable JSON
- Agent Response — JSON Schema
- LLM content map — optional machine-readable site summary
Challenge v1.2.0 · MIQP v1.0.0 · Local-browser self-assessment only.