OpenAI 2026 hackathon

ReplayLab

ReplayLab turns captured user sessions into validated GPT-5.6 reproduction plans, then replays the exact experience and reproduces the bug safely in Playwright.

Team of 3 · 0 likes · 0 comments

Archive position — measured, not model output

0 likes on Devpost

2,264 of the 7,856 archived projects have more likes, and 5,592 share exactly 0 — so this project's #6,354 place in the like-ranked listing is a tie-break inside that group, not a ranking.

Projects (log scale)

1
10
100
1k
10k
05,592
11,758
2285
3–4132
5–975
10+14

Likes on Devpost. ▲ marks this project's group.

Show the figures
LikesProjectsShare of archive
05,59271.2%
11,75822.4%
22853.6%
3–41321.7%
5–9751.0%
10+140.2%
Devpost like counts for all 7,856 archived projects, captured when this archive was built.

Executive Summary

What the company appears to be

ReplayLab is a self-reported developer tool that captures user sessions in web applications, replays them visually, and uses GPT-5.6 to generate structured reproduction plans for bugs. These plans are validated and then executed safely using Playwright against a local test environment.

What changed

The project description shows an evolution from a general idea ("turn customer support tickets into executable workflows") to a specific technical implementation involving session capture, AI planning with constraints, and deterministic execution via Playwright.

Single most important open question

Is there any evidence of real-world usage or traction beyond the hackathon demo? The description states no revenue, customers, or adoption data exist outside of the author's own account.

Back to contents

What The Product Actually Is

The description states that ReplayLab is:

  • An Electron desktop application built with React and TypeScript
  • A system for capturing user sessions in web apps using rrweb
  • A tool that replays captured sessions visually with controls like play, pause, seek
  • A platform that integrates GPT-5.6 to convert evidence into structured reproduction plans
  • A system that safely reproduces bugs using Playwright against local fixtures

The product is described as a four-step workflow:

  1. Capture user session (DOM changes, actions, state snapshots)
  2. Replay exact experience visually
  3. Generate reproduction plan with GPT-5.6
  4. Reproduce bug safely in Playwright

Back to contents

Positioning & Claim Evolution

The description states that the company's positioning evolved from:

  • Initial claim: "turn customer support tickets into executable reproduction workflow"
  • Specific implementation: "capture what user actually experienced, replay session visually, ask GPT to convert evidence into constrained reproduction plan"

Claims made:

  • The tool addresses a problem where support tickets lack context for developers
  • It turns vague tickets into executable workflows
  • It uses AI between evidence and deterministic execution
  • It prevents arbitrary AI output by using allowlists and structured outputs

Back to contents

Target Customer & ICP

The description states that the target customer is:

  • Developers who need to reproduce bugs from customer support tickets
  • Teams working with web applications where user sessions are captured
  • Organizations that want to improve handoffs between support and development teams

No specific customer segments or personas are named. The positioning appears to be for technical users in software development environments.

Back to contents

Business Model & Pricing Evidence

Not evidenced. The description does not contain any information about pricing, monetization, or business model.

Back to contents

Technical & Delivery Signals

The description states that ReplayLab uses:

  • Electron desktop application with React and TypeScript
  • rrweb for DOM recording
  • Playwright for browser automation
  • GPT-5.6 via OpenAI Responses API
  • Structured Outputs with Zod validation
  • Security boundaries including encryption of API keys
  • In-memory dry-run before execution
  • Fixed adapters to prevent arbitrary code generation

The system is described as having:

  • Multiple security layers
  • Strict validation at each step
  • Deterministic execution paths
  • Local-only execution environment
  • Exportable .replaylab bundles

Back to contents

Traction & Maturity Signals

Not evidenced. The description states that this was submitted to an OpenAI hackathon and contains no information about revenue, customers, usage metrics, or adoption beyond the demo.

Back to contents

Competitive Context

Not evidenced. No mention of competitors or market positioning beyond what the authors describe.

Back to contents

Key Risks & Red Flags

Inferences based on the self-reported description:

  • The tool is described as an MVP for a Demo Store environment only
  • No evidence of support for arbitrary production websites
  • The system relies heavily on GPT-5.6 which may not be available in all contexts
  • The entire workflow requires user input and manual export/import steps
  • No information about scalability or enterprise readiness

Back to contents

Diligence Questions To Ask The Founders

  1. What is the actual use case beyond the hackathon demo?
  2. How does this integrate with existing development workflows?
  3. Are there any customers or pilot users currently using this?
  4. What are the technical limitations of the current MVP that would prevent production deployment?
  5. How does the team plan to scale beyond the current demo environment?
  6. What is the roadmap for supporting third-party web applications?
  7. How do you handle edge cases in session capture and replay?
  8. What happens when GPT-5.6 fails to generate a valid plan?

Back to contents

Investment/Partnership Verdict

Not evidenced. The description contains no information about funding, valuation, or investment status beyond the fact that it was submitted to a hackathon. No commercial traction or financial data is provided.

The product appears to be a proof-of-concept for a specific use case (bug reproduction in controlled environments) rather than a mature commercial offering. The self-reported description lacks evidence of real-world adoption or business metrics.

Back to contents

Source

Submitted to the OpenAI 2026 hackathon on Devpost. Project home on DevPost.

The analysis above was generated by a language model from the project's own one-line description. It is not independent research and contains no verified traction, revenue or customer data.