OpenAI 2026 hackathon

Judgment Mirror

AI outputs are observable while judgment shifts during AI interactions usually aren't. Judgment Mirror makes them visible.

Solo project by biasbreaker Myers · 0 likes · 0 comments

Archive position — measured, not model output

0 likes on Devpost

2,264 of the 7,856 archived projects have more likes, and 5,592 share exactly 0 — so this project's #4,738 place in the like-ranked listing is a tie-break inside that group, not a ranking.

Projects (log scale)

1
10
100
1k
10k
05,592
11,758
2285
3–4132
5–975
10+14

Likes on Devpost. ▲ marks this project's group.

Show the figures
LikesProjectsShare of archive
05,59271.2%
11,75822.4%
22853.6%
3–41321.7%
5–9751.0%
10+140.2%
Devpost like counts for all 7,856 archived projects, captured when this archive was built.

Executive Summary

What the company appears to be

Judgment Mirror is an experimental observatory for human judgment during AI interaction, built as a prototype for the OpenAI 2026 hackathon. It guides users through a decision-making scenario with AI assistance and visualizes changes in their judgment before and after AI input.

What changed

The project description indicates this is a self-contained prototype built by one person (biasbreaker Myers) using Next.js, TypeScript, GPT-5.6, and Codex. It does not appear to have moved beyond the prototype stage or demonstrated any commercial traction.

Single most important open question

Is there evidence of a viable market need for human-AI judgment observability that would justify further development beyond this experimental prototype?

The description states that Judgment Mirror is an experimental observatory, not a product with revenue, customers, or adoption. It was built as a hackathon submission and contains no evidence of commercial viability, funding, or user base.

Back to contents

What The Product Actually Is

The description states that Judgment Mirror is:

  • An experimental observatory for human judgment during AI interaction
  • A standalone Next.js and TypeScript application
  • Built with GPT-5.6 via the OpenAI Responses API
  • Designed to guide users through realistic operational decision scenarios
  • Not an intelligence test, psychological assessment, truth detector, or decision-grading system
  • Not designed to determine whether the user or AI is correct

The product does not appear to have moved beyond a prototype stage. It stores complete sessions only in browser state and contains no authentication, database, user profile, payment system, uploads, or unrelated features.

Back to contents

Positioning & Claim Evolution

The description states that Judgment Mirror:

  • Focuses on making invisible judgment shifts visible during AI interactions
  • Started with the question: "What changed in a person's decision process after the AI answered?"
  • Is not an intelligence test, psychological assessment, truth detector, or decision-grading system
  • Does not determine whether the user or AI is correct
  • Visualizes the decision process that occurred during one interaction

The positioning appears to be:

  • An experimental tool for observing human-AI interaction dynamics
  • Focused on transparency in judgment processes rather than evaluation of correctness
  • Positioned as a research instrument, not a commercial product

Back to contents

Target Customer & ICP

The description states that Judgment Mirror:

  • Guides users through one realistic operational decision scenario
  • Is designed to observe the information recorded during one AI-assisted decision process
  • Does not claim to evaluate intelligence, reasoning ability, or diagnose cognition/mental health

No specific target customer segment is identified. The description indicates it's an experimental tool for observing human-AI interactions, but does not specify who would use it or what their role might be.

Back to contents

Business Model & Pricing Evidence

The description states:

  • Judgment Mirror is an experimental observatory
  • It was built as a prototype for the OpenAI 2026 hackathon
  • The prototype stores complete sessions only in browser state
  • It contains no authentication, database, user profile, payment system, uploads, or unrelated features
  • No pricing information or business model is provided

There is no evidence of any commercial business model, pricing structure, or monetization strategy.

Back to contents

Technical & Delivery Signals

The description states that Judgment Mirror:

  • Is built as a standalone Next.js and TypeScript application
  • Uses GPT-5.6 via the OpenAI Responses API
  • Was implemented with Codex as an implementation partner
  • Uses structured JSON outputs validated through strict schemas (Zod)
  • Performs two bounded tasks: generating AI recommendation and analyzing visible differences between recorded judgment stages
  • Does not control user snapshots, decision comparisons, evidence-coverage arithmetic, metric calculations, or stage attribution through the model
  • Enforces these through deterministic application logic
  • Stores complete session only in browser state
  • Contains no authentication, database, user profile, payment system, uploads, or unrelated features

The technical implementation shows:

  • Use of modern web stack (Next.js, TypeScript)
  • Integration with OpenAI APIs
  • Structured data validation
  • Deterministic logic for integrity controls
  • Automated testing (80 automated tests)

Back to contents

Traction & Maturity Signals

The description states:

  • Judgment Mirror is an experimental observatory
  • It was built as a prototype for the OpenAI 2026 hackathon
  • The prototype passed 80 automated tests, TypeScript validation, ESLint with zero warnings, production build validation, original failure-pattern integration fixture, and three complete live Scenario 01 journeys
  • All three live journeys rendered one complete Judgment Mirror without manual retry
  • Every run preserved all four decision states independently, retained AI-generated provenance, used neutral reflection language, and correctly handled the no-claim evidence state

No evidence of revenue, customers, or adoption beyond the prototype testing phase.

Back to contents

Competitive Context

The description states:

  • Judgment Mirror is experimental and not scientifically calibrated
  • It does not evaluate intelligence, reasoning ability, diagnose cognition or mental health, determine whether a decision is correct, or replace professional decision systems
  • It observes only information recorded during one AI-assisted decision process

No competitive landscape or market positioning relative to other tools is described. The description indicates this is an experimental tool without established competitors.

Back to contents

Key Risks & Red Flags

The description states:

  • Judgment Mirror is experimental and not scientifically calibrated
  • It does not evaluate intelligence, reasoning ability, diagnose cognition or mental health, determine whether a decision is correct, or replace professional decision systems
  • It observes only information recorded during one AI-assisted decision process
  • The next stage is not to add more participant-facing metrics and call them science
  • Controlled pilot studies, versioned scenarios, preregistered hypotheses, reliability testing, construct validation, independent replication, and clearly bounded scientific claims are required before it could become a validated research instrument

Key risks include:

  • Prototype-only status with no commercial traction
  • Experimental nature without scientific calibration
  • No evidence of market demand or user adoption
  • No clear path to commercialization

Back to contents

Diligence Questions To Ask The Founders

  1. What specific problem are you trying to solve that would justify moving beyond this experimental prototype?
  2. Have you identified any potential users or customers who would pay for this functionality?
  3. What evidence do you have of market demand for human-AI judgment observability?
  4. How would you transition from this prototype to a commercial product?
  5. What are the key metrics that would indicate success beyond the current experimental phase?
  6. Have you considered how to validate the scientific claims and ensure reproducibility?
  7. What is your plan for scaling beyond the single-person development team?

Back to contents

Investment/Partnership Verdict

The description states that Judgment Mirror:

  • Is an experimental observatory for human judgment during AI interaction
  • Was built as a prototype for the OpenAI 2026 hackathon
  • Contains no evidence of revenue, customers, or adoption beyond the prototype testing phase
  • Does not evaluate intelligence, reasoning ability, diagnose cognition or mental health, determine whether a decision is correct, or replace professional decision systems

Verdict Not evidenced. The description shows this is an experimental prototype with no commercial traction, revenue, customers, or demonstrated market need. It lacks any evidence of a viable business model, target customer base, or path to monetization. The project appears to be a proof-of-concept rather than a commercial venture.

Back to contents

Source

Submitted to the OpenAI 2026 hackathon on Devpost. Project home on DevPost.

The analysis above was generated by a language model from the project's own one-line description. It is not independent research and contains no verified traction, revenue or customer data.