Expert verdicts on what AI builds.
World Labs Community

The Panel

Less vibes. More verdicts.

Play and judge what AI builds. Get paid for the judgment you spent a career earning.

Five specialties. One bar: shipped.

Every panelist has shipped titles behind them.

Level Design
Systems Design
UX & Onboarding
QA & Playtest Leads
Game Feel

No exclusivity requirement, no middleman markup. Just a curated panel of shipped game developers who share a high bar for what counts as a good game.

Three steps in. Then you're in.

1

Apply

Send your shipped credits and specialty. Fifteen minutes, no portfolio required.

2

Calibrate

A short conversation and one sample eval. We're checking judgment, not trivia.

3

You're in

Assignments start with the next round. No probation tiers, no ranking games.

More than a gig queue

Paid work

Eval sessions matched to your expertise

Assignments route by specialty. Paid per completed eval, invoiced monthly.

Community

A private room of people who ship

A members-only space to compare notes and see where AI game tools are actually heading.

Early access

The tools, before the hype

Play the newest prompt-to-game tools under blind conditions, first.

On your terms

Flexible, remote, no exclusivity

Take the rounds you want, skip the rest. Everything runs in the browser, on your schedule.

Fair questions, straight answers

What does the work actually look like?
You're assigned a game generated by an AI tool (identity hidden), you play it in the browser for a minimum session, usually around 15 minutes, then score it on a seven-dimension rubric, write a critique, and rank the top three fixes. Some assignments are head-to-head A/B comparisons instead.
How does payment work?
Flat rate per completed eval, set per round before you accept any assignments. A rubric session pays more than an A/B comparison. No bidding, no rate negotiation per gig, no platform cut ambiguity. Invoices settle monthly.
What's the time commitment?
Whatever you take on. A typical round assignment is a handful of sessions over two to three weeks, each about 30 to 40 minutes including the writeup. You can decline any round with zero penalty. The only expectation is that accepted assignments get finished.
Do I need AI experience?
No. We're paying for the opposite: judgment about games, built by shipping them. If you can tell a coherent mechanic from a broken one and explain why in writing, you're qualified. The harness handles everything else.
Is my name attached to my scores?
Never. Published results and licensed data are attributed only to specialty and years of experience: "Level designer, 12 years," not your name. Your credits get you in the door; your identity stays out of the data.
How is this different from a playtesting marketplace or staffing agency?
Marketplaces sell volume; we sell calibrated expert judgment. Every panelist is vetted on shipped titles, every eval follows the same anchored rubric, and every game gets multiple independent raters with disagreements flagged and re-judged. Agencies place you with clients. We run the whole harness and you just judge.

Your judgment is the product.