import { benchmarkDecisionPolicy, decisionState, modelTargets, multimodalBenchmarkSamples, productionNorthStars, } from "../../../lib/model-performance"; export default function ModelEconomicsPage() { const imageModels = modelTargets.filter((model) => model.modality === "image"); const videoModels = modelTargets.filter((model) => model.modality === "video"); const voiceModels = modelTargets.filter((model) => model.modality === "voice"); return ( <>
Production intelligence

Image · Video · Voice Model Performance

CartoonOS compares matched production tasks by quality, identity consistency, first-pass acceptance, latency, rework and accepted-output cost. There is no single global best model.

{multimodalBenchmarkSamples.length} matched samples
Image economics

{productionNorthStars.imageEconomics}

Character sheets, turnarounds, expressions, clean backgrounds, diagrams, storyboards and keyframes.

Video economics

{productionNorthStars.videoEconomics}

Dialogue, cinematic shots, multi-character scenes, B-roll, Clarity Islands and motion transfer.

Voice economics

{productionNorthStars.voiceEconomics}

Canonical mascot voice identity is measured separately from the underlying TTS engine.

Current benchmark portfolio

{imageModels.length} image · {videoModels.length} video · {voiceModels.length} voice
{modelTargets.map((model) => { const samples = multimodalBenchmarkSamples.filter( (sample) => sample.providerId === model.providerId && sample.modelId === model.modelId, ); return ( ); })}
Modality Provider Model Route Unit Samples Decision state
{model.modality} {model.providerName} {model.modelName}
{model.modelId}
{model.route} {model.benchmarkUnit.replaceAll("_", " ")} {samples.length} {decisionState(samples.length).replaceAll("_", " ")}
Image focus

Nano Banana is in the benchmark set

Nano Banana, Pro, 2 and 2 Lite are evaluated against the same character/reference briefs. Price alone never overrides identity or anatomy gates.

Voice law

Same mascot voice across formats

Lumi, Tiko, Nova, Marina, Rin, Arqueo and Pepe each keep one active versioned voice identity across Shorts/Reels and flagships.

Confidence

No winner from tiny samples

<5 matched samples = insufficient; 5–19 = provisional; ≥{benchmarkDecisionPolicy.decisionReadyMin} = decision-ready.

Quality KPI

First-pass acceptance

Measures how often a model clears the exact quality gate without regeneration or repair.

Identity KPI

Character / voice consistency

Character assets and dialogue must preserve the approved identity before economics may influence routing.

Speed KPI

P50 / P95 latency

Queue, generation, QA and total wall-clock time are tracked separately to expose scale bottlenecks.

Selector rule

Best model for this task

Rank by matched archetype, exact character/version, quality floor, expected accepted cost, reliability and latency.

Portfolio rule

No global best model

A cheap background model can be a poor Lumi close-up model; a strong image model may not be the best storyboard or diagram model.

North Star

{productionNorthStars.primary}

Production-model optimization ultimately serves durable audience value, not API-price minimization.

); }