The independent infrastructure
for pricing, verifying, and
settling autonomous work.

  • Agent.so

    Anything

    Bubble

    Buildship

    Flora

    FlutterFlow

    Framer

    Framer Commerce

    Frameship

    General Assembly

    Glide

    Glorify

    HeyGen

    Ideogram

    Instant

    Jitter

    Kajabi

    Kit (ConvertKit)

    Kittl

    LottieFiles

    Lovable

    Lovart

    Lummi

    MagicPath

    Pagedeck

    PeachWeb

    ReadyMag

    Relume

    Replit

    Replo

    Retool

    Rive

    Spline

    Stripo

    Tempo Labs

    Unicorn Studio

    Webflow

    Webstudio

  • AI agents

    AI 3D generation

    No-code apps

    Backend automation

    Creative collaboration

    Mobile app builder

    Website builder

    Ecommerce platform

    Developer infrastructure

    Tech education

    No-code apps

    AI design tool

    AI video generation

    AI image generation

    AI eCommerce builder

    Motion design

    Course platform

    Email marketing

    AI graphic design

    Animation platform

    AI app builder

    AI art generation

    AI image platform

    AI workflows

    Presentation builder

    Website builder

    Website builder

    AI wireframing

    AI coding platform

    Ecommerce builder

    Internal tools builder

    Interactive animation

    3D design

    Email design

    AI development

    Interactive design

    Website builder

    Website builder

Different models. Different markets.
One contract for completed work.

Learn about [Ontitled]
ecosystem

We build the eval layer for creative AI, helping AI tools and teams understand quality the way expert creatives do.

Public launch-focused research

Creative

Arena

AI models go head to head on creative output. Real creatives judge the results. The best model wins. Human taste, applied.

Public launch-focused research

Creative

Arena

AI models go head to head on creative output. Real creatives judge the results. The best model wins. Human taste, applied.

Public launch-focused research

Creative

Arena

AI models go head to head on creative output. Real creatives judge the results. The best model wins. Human taste, applied.

Comparative model evals

Benchmarking

The industry standard for measuring creative AI. Built on the judgments of 1.5M+ creatives, HCB answers the only question that matters: is this actually good?

Comparative model evals

Benchmarking

The industry standard for measuring creative AI. Built on the judgments of 1.5M+ creatives, HCB answers the only question that matters: is this actually good?

Comparative model evals

Benchmarking

The industry standard for measuring creative AI. Built on the judgments of 1.5M+ creatives, HCB answers the only question that matters: is this actually good?

Private lab benchmarks

Creative

preference

Seed datasets (image, video, UI, motion). Corrective SFT data and RLHF preference pairs.

Private lab benchmarks

Creative

preference

Seed datasets (image, video, UI, motion). Corrective SFT data and RLHF preference pairs.

Private lab benchmarks

Creative

preference

Seed datasets (image, video, UI, motion). Corrective SFT data and RLHF preference pairs.

Long-horizon workflow capture

Trajectories

Computer-use trajectories from real-world projects with reasoning

Long-horizon workflow capture

Trajectories

Computer-use trajectories from real-world projects with reasoning

Long-horizon workflow capture

Trajectories

Computer-use trajectories from real-world projects with reasoning

Public launch-focused research

Creative

Arena

AI models go head to head on creative output. Real creatives judge the results. The best model wins. Human taste, applied.

Comparative model evals

Benchmarking

The industry standard for measuring creative AI. Built on the judgments of 1.5M+ creatives, HCB answers the only question that matters: is this actually good?

Private lab benchmarks

Creative

preference

Seed datasets (image, video, UI, motion). Corrective SFT data and RLHF preference pairs.

Long-horizon workflow capture

Trajectories

Computer-use trajectories from real-world projects with reasoning

Building taste into creative AI

Ontitled brings together private evaluations, public comparisons, and research outputs under one unified human evaluation framework.

Blind comparative reviews

Models are evaluated head-to-head through structured, bias-controlled comparison protocols.

Expert preference modeling

We capture and analyze high-signal human preferences from vetted creative professionals.

Creative scoring systems

Qualitative judgment is translated into structured scoring frameworks aligned with real-world standards.

Multimodal AI evals

We assess image, video, design, and interactive model outputs where traditional metrics fail to capture taste.

Built on real-world
creative expertise

The

The

Network

Network

of Human

of Human

Taste

Taste

1.5M+

independent creatives

1.5M+

independent creatives

400+

different creative skills

400+

different creative skills

$250M+

earned by creatives

$250M+

earned by creatives

26x

higher project earnings by creatives on [Ontitled]

26x

higher project earnings by creatives on [Ontitled]

50+

models evaluated

50+

models evaluated

Work with the industry’s top creative minds

Work with the industry’s top creative minds

Work with the industry’s top creative minds

Work with the industry’s top creative minds

The creative class sets the standard.

Leading next-gen creative AI research.

Human taste is the new training data.

Creative AI isn't good enough yet. Different models. Different markets. One contract for completed work.

Execution is abundant.

Execution is abundant.

Verified completion is scarce.

Stop paying for unmetered token consumption and failed agent loops. Ontitled calculates risk upfront and settles transactions only when your system of record confirms the work is done.