Why one number is not enough
A single AI score for an ad tells you nothing actionable. It does not explain what is wrong, where to look, or what to fix. The score might be 72, but is that good? Compared to what? And if it is bad, which part of the creative is causing the problem?
RoastIQ, the scoring engine behind SaliencyLab's pre-spend creative intelligence platform, solves this with three layers of scoring: a verdict before media budget is committed. Each layer answers a different question at a different altitude.
Layer 1: Raw perception signals
The bottom layer estimates five atomic perception signals:
- Attention (0-100): How effectively the creative captures and holds visual attention
- Clarity (0-100): How clearly the message and value proposition communicate
- Branding (0-100): How strongly the brand registers, including logo timing, brand colors, and recall cues
- Emotion (0-100): The emotional response the creative generates
- CTA (0-100): How effectively the call to action drives intended behavior
These are the raw inputs. They come from the multimodal analysis pipeline: visual attention prediction, transcript scoring, brand signal detection, and audio analysis.
Layer 2: Sub-KPI families
The middle layer combines raw signals into diagnostic families:
Get Noticed components: Branding, Enjoyment (Emotion), and Average Ad Viewed (Attention)
Sell Proposition components: Persuasion, Clarity, and CTA
Build Brand components: Meaningful, Different, and Salient
This layer tells you not just that Get Noticed is weak, but which component is dragging it down. Is it the branding, the emotional response, or the view depth?
Layer 3: The 5 main KPIs and composite
The top layer produces the five KPI scores that determine the verdict:
- Beat the Skip (25%): Attention (55%) plus skip retention (45%)
- Get Noticed (20%): Branding (60%) plus Emotion (40%)
- Brand Impact (20%): Branding (45%) plus brand lift (30%) plus Emotion (25%)
- Sell Proposition (20%): Conversion (60%) plus CTA (40%)
- Build Brand (15%): Brand lift (55%) plus Branding (30%) plus Clarity (15%)
The weighted composite then maps to one of three verdicts.
How the verdict is decided
The three verdict states are Scale, Sharpen, and Rebuild. The rules are strict, and the per-KPI floors matter as much as the composite:
- Scale: composite of 70 or above, and no individual KPI below 55
- Sharpen: composite between 55 and 69. Also any creative scoring 70 or above that has a KPI below 55, because the floor rule blocks it from Scale
- Rebuild: composite below 55, or two or more KPIs below 45
That second rule catches a case teams often miss. A creative can post a composite above 70 and still not earn a Scale verdict, because one weak KPI drags it into Sharpen. A strong average with one structural hole is not the same thing as a creative that is ready to run.
One caveat applies to every layer: these scores are model predictions, not measurements. In held-out, out-of-sample validation as of May 2026 they correlate +0.30 to +0.32 (Spearman ρ) with public engagement and click-intent outcomes, not with ROAS, sales, or brand recall. The full record is on the methodology page.
Why three layers matter
A creative director who sees Sharpen with Sell Proposition at 63 can drill into Layer 2 and find that Clarity is strong but CTA is weak. Then drill into Layer 1 and see the raw CTA score. Each layer narrows the diagnosis until the fix becomes specific. For what each KPI is asking of the creative, see how to interpret RoastIQ KPI scores.
One number hides the problem. Three layers reveal it.

