Evidence-based AI ad optimization

The AI ad optimizer that shows its work.

Connect your Meta and Google Ads accounts. Mervin compares each proposed action against real outcomes from similar campaigns, then shows the sample size, median effect, confidence interval, and its own historical hit rate before you approve anything.

Every action requires approvalNew ads start pausedMeta + Google AdsOfficial Marketing APIs
Connected · 12 accounts
YOU · 09:42

Should I increase the budget for this campaign?

MERVIN · 09:42

I compared this campaign with similar budget changes from relevant historical cases. Here is the evidence before you decide.

RECOMMENDATION

Increase daily budget by 15%

Illustrative demo data
CurrentNT$3,000/day
ProposedNT$3,450/day
Samplen=186
Median ROAS+8.4%
Positive67%
95% CI+2.1 to +13.7%
-5%0+10%+20%
Statistically significant · 14-day window

Similar actions were followed by a positive ROAS change in 67% of observations. Historical association is not a guarantee.

EVERY RECOMMENDATIONCOMES WITH EVIDENCEAND WAITS FOR APPROVAL

The recommendation, fully unpacked

One proposal. Every reason visible.

See the action, cohort definition, observed outcomes, uncertainty, limitations, and the exact write operation before making a decision.

Illustrative demo dataModerate risk

RECOMMENDATION · META ADS

Increase daily budget by 15%

TW Commerce / Meta Ads · Q4 Prospecting / Broad Value

Active / stable delivery
Current budgetNT$3,000/day
Proposed budgetNT$3,450/day+NT$450 / +15%
WHY NOW

ROAS has remained above the account target for 7 days while impression share is constrained by budget.

EVIDENCE FROM SIMILAR ACTIONS

What happened after comparable changes

Significant
Sample size186observations
Median 14-day ROAS+8.4%historical change
Positive outcomes67%125 / 186
95% confidence interval+2.1% → +13.7%does not cross zero
14-day ROAS change95% CONFIDENCE INTERVAL
+8.4%
-10%0%+10%+20%
PLAIN-LANGUAGE INTERPRETATION

In similar historical cases, this type of budget increase was followed by a positive ROAS change in 67% of observations. The result is statistically meaningful, but it does not guarantee the same outcome for this campaign.

Limitations & exact changes
LIMITATIONS

Historical association does not establish causation. Auction pressure, creative fatigue, or tracking changes may produce a different result.

EXACT CHANGES PREVIEW
  • Set daily_budget from 300000 to 345000 (minor currency units)
  • Keep campaign status ACTIVE
  • Start a 14-day observation window after API confirmation

Nothing changes until you approve.

The black-box problem

Most AI tools give you an answer. Mervin gives you the evidence.

Typical AI recommendation
Increase the budget by 20%.
  • No sample size
  • No historical outcome
  • No confidence interval
  • No accuracy record
  • No comparable-case definition
Mervin recommendation
In 186 comparable cases, similar budget changes were followed by a median 8.4% ROAS increase, with a 67% positive outcome rate.
n=18695% CI +2.1 to +13.7%Significant

Not “the AI thinks.” The data shows.

Evidence Engine

Change the cohort. Watch the evidence change.

Every condition remains inspectable. Adjust the platform or industry to see when the historical signal is meaningful, and when Mervin should say there is not enough evidence.

Comparable cohort8 filters

Active cohort: Meta Ads, E-commerce, Conversions, Taiwan, Q4, budget increase of 10–20%, evaluated over 14 days.

COHORT RESULT · ILLUSTRATIVE DEMO DATA

Evidence passes the display threshold

Significant
Sample size186observations
Median change+8.4%ROAS / 14 days
Positive67%outcomes
UpdatedDailyToday, 06:10 UTC+8
Outcome distributionROAS · %
-20-100+10+20+30
Outcome composition186 observations
67%14%19%
Positive Neutral Negative
14-day ROAS change95% CONFIDENCE INTERVAL
+8.4%
-10%0%+10%+20%

For this cohort, budget increases were historically followed by a median 8.4% ROAS change over 14 days. 67% of observed outcomes were positive. Historical evidence is not a guaranteed forecast.

How Mervin learns

Every approved action makes the next recommendation more accountable.

Mervin turns actions and subsequent observations into a decision record that can be reviewed, challenged, and scored.

01

Connects

Meta and Google Ads accounts

02

Reads

Account and performance history

03

Records

Historical advertising actions

04

Observes

Performance after each action

05

Matches

Relevant comparable cases

06

Proposes

An evidence-backed action

07

Scores

The result and its own hit rate

The statistical engine is designed to learn from dozens of ad accounts, hundreds of thousands of daily performance records, and more than ten thousand historical actions, with daily updates. These are scale ranges, not exact performance claims.

A day with Mervin

Decision support that follows the working day.

Mervin reads continuously, but the human stays at the decision point.

01

Mervin briefs you

Yesterday's Spend, ROAS, CPA, conversions, and unusual changes are summarized across accounts.

02

A decision point appears

Mervin flags a budget, creative, keyword, audience, or performance issue that needs judgment.

03

Mervin checks the evidence

Comparable actions are matched and evaluated for sample size, median effect, positive rate, confidence, and significance.

04

Mervin executes

The exact proposal is sent through the official API, logged, and monitored only after a person approves it.

Approval mode is always on. Mervin never changes a campaign without permission.

Mervin OS

Your accounts, recommendations, evidence, and approvals in one place.

A work-focused command center built around the question that matters: why is this action worth considering?

WORKSPACE / RECOMMENDATIONS

Good morning, Wayne.

SpendNT$128.4K+3.2% vs prev. 7d
RevenueNT$492.7K+8.6% vs prev. 7d
ROAS3.84×+0.18 vs prev. 7d
CPANT$612-4.1% vs prev. 7d
Conversions805+6.7% vs prev. 7d
ACCOUNT HEALTH

Performance overview

Healthy · 86
150K100K50K0
M
T
W
T
F
S
S
Revenue SpendLast 7 days · NT$
NEEDS ATTENTION

Performance anomalies

3 open
  • CPA increased 18%Google · Brand Search · 3h ago
  • ROAS above targetMeta · Broad Value · 6h ago
  • Creative fatigue signalMeta · Retargeting · 1d ago
WHY MERVIN RECOMMENDS THIS

Increase Broad Value budget by 15%

Illustrative demo data

Spend is budget-constrained while 7-day ROAS remains above the account target. The matching cohort excludes learning-phase campaigns.

Sample186Median ROAS+8.4%Positive67%95% CI+2.1 → +13.7%
ACTION QUEUE

Approvals & monitoring

4Pending approval
7Monitoring
72%Hit rate
Budget increase · Broad ValueAwaiting approval · 12m
Pause · Low CTR creativeMonitoring outcome · Day 5 of 14

Transparent accountability

Mervin keeps score on itself.

A recommendation is not counted as correct or incorrect until its observation window closes. Pending outcomes never inflate the score.

Illustrative demo data

OVERALL HIT RATE

71.2%
71.2%

Correct outcomes among recommendations whose 14-day evaluation window has closed.

Correct
146
Incorrect
59
Pending
34
Evaluated sample
n=205
Last updated today, 06:10 UTC+8
BY RECOMMENDATION TYPE

Where the record is strong, weak, or still forming

14-day evaluation
Budget increasen=68 closed recommendations
72%
Budget decreasen=42 closed recommendations
69%
Pause recommendationn=37 closed recommendations
76%
Keyword adjustmentn=31 closed recommendations
64%
Audience changen=28 closed recommendations
61%
New campaignn=9 closed recommendations
Not available

What counts as a hit? The recommendation's stated outcome must beat its pre-defined baseline after the full evaluation period. Recommendations still inside that period are pending and excluded from the hit rate.

Approval-first execution

See the evidence. Approve the move.

Inspect the exact operation, adjust it if needed, then approve. Try the demo below to move a proposal from draft to execution monitoring.

ACTION PROPOSAL · ILLUSTRATIVE DEMO DATA

Increase campaign daily budget

Draft
TW CommerceQ4 ProspectingBroad Value
Current settingNT$3,000/day
Proposed settingNT$3,450/day
Expected outcomeHistorically associated with median +8.4% ROAS
Supporting evidencen=186 · 67% positive · significant
RiskModerate · auction response may differ
Exact API operationUpdate campaign daily_budget only
Requested by
Wayne · Performance Lead
Created
Today, 14:20 UTC+8
New-ad default
Paused until separately approved

Core capabilities

Evidence is not a feature. It is the operating model.

01

Conversational operation

Discuss an account like you would with a senior optimizer: inspect numbers, test reasoning, and adjust the direction.

02

Evidence-backed recommendations

Every proposal includes sample size, median effect, confidence interval, and statistical significance.

03

Recommendation accuracy tracking

Review how past recommendations performed after their observation windows closed.

04

Approval-first execution

Every write action needs human approval. Newly created ads start paused.

05

Meta and Google Ads

Analyze and manage both advertising platforms from one consistent decision workspace.

Who it is for

Built for teams that need judgment to scale.

Performance marketers

Review more accounts without giving up professional judgment. See where attention is needed and why.

Explore Mervin

Agencies

Turn senior optimization experience into a reviewable decision record that teams can discuss, approve, and hand over.

Explore Mervin

Advertisers

Move beyond 'the AI recommends.' Inspect the evidence and decide whether the action is worth taking.

Explore Mervin

A clearer decision system

Experience matters. Evidence makes it inspectable.

Mervin combines human judgment with a visible statistical record and a controlled execution path.

LIMITED VISIBILITY

Human intuition only

  • Depends on individual experience
  • Difficult to quantify
  • Hard to transfer
  • Limited by personal memory
LIMITED VISIBILITY

Black-box AI tools

  • Answers without evidence
  • No visible sample size
  • No accuracy record
  • Unclear why actions occur

Transparent by default

Questions the evidence should answer.

No “proprietary AI” shortcut. Just clear definitions, limits, and operating rules.

Ask about your accounts
What makes Mervin different from other AI optimization tools?

Mervin does not stop at a recommendation. It shows the comparable cohort, sample size, historical outcome distribution, confidence interval, limitations, and Mervin's track record for that recommendation type before any action is approved.

What counts as a similar advertising case?

Similarity is defined with visible filters such as platform, industry, objective, market, season, action type, change range, tracking quality, and campaign state. The cohort definition appears on every recommendation so a reviewer can challenge it.

Where does the historical evidence come from?

The engine learns from advertising actions and subsequent performance observations available to the product. Results are presented as aggregated cohorts. This page uses illustrative demo data, not customer performance claims.

Does statistical significance guarantee better performance?

No. Significance means the observed pattern passed a defined statistical threshold in the historical cohort. It does not prove causation or guarantee that a future campaign will behave the same way.

How is Mervin's recommendation hit rate calculated?

A recommendation is evaluated only after its observation window closes. Its outcome is compared with the stated baseline and success definition. Pending recommendations are excluded from the rate, and the sample size is always shown.

What happens when there is not enough data?

Mervin labels the proposal 'Insufficient evidence,' explains which cohort is too small, and avoids presenting statistical significance. The user can still inspect the account context and decide manually.

Does Mervin expose another advertiser's account data?

Recommendation evidence is designed to show aggregated cohort results, not another advertiser's campaign names, creatives, audiences, or raw account records. The exact production safeguards should be confirmed during security review.

Can Mervin change campaigns without approval?

No. Approval mode is always on. Budget changes, pauses, audience edits, and newly created ads all remain proposals until a person approves the exact operation.

Why are newly created ads paused by default?

A paused default prevents a new ad from spending immediately after creation. A reviewer can inspect the creative, targeting, budget, and tracking setup before activation.

Which Meta and Google Ads actions are supported?

The interface is designed for budget, status, audience, keyword, bid, and new-ad proposals. The exact supported write operations depend on the connected platform and the current product release.

How frequently is the evidence engine updated?

The engine is designed to refresh daily as new performance observations complete their evaluation windows. Each evidence panel shows its own data freshness timestamp.

Can agencies manage multiple accounts?

Yes. Mervin OS is designed around an account selector, shared approval queue, action history, and evidence records so agency teams can review decisions across accounts without losing context.

Every recommendation comes with evidence.

Your next optimization decision should come with evidence.

Connect an account, review your first evidence-backed recommendation, and decide whether it is worth acting on.

No automatic changesEvery action requires approvalNew ads start pausedConnected through official APIs