Original research protocol

I Framed the Creator Panic-Response Study Around Evidence, Not Algorithms

A proposed study of what creators change after an underperforming video, designed to compare evidence without treating ViralJury as ground truth.

A video falls far below a creator's usual range. By lunch, the creator has changed the next hook, posting time, hashtag set, niche, and upload schedule. The urgency is real. The diagnosis may still be unknowable.

This study is not built to laugh at that reaction. Creators make business and identity decisions inside systems they cannot fully observe. Native analytics describe audience behavior after publication but do not expose a complete causal model. A structural review can identify controllable features in the file but cannot prove why distribution stopped.

The research question is therefore about evidence alignment, not right and wrong. What did the creator believe? What did they do first? What evidence could they access at that moment? Did the action address that evidence, partly address it, target something else, or remain impossible to judge?

A flop is relative to a declared baseline

The study should not define every underperforming video as 40 percent below expected views at 48 hours. Platforms, accounts, formats, and normal volatility differ too much for one universal cutoff. Each creator supplies a registered comparison set of recent videos similar in format and duration, with the median and spread of the available metrics.

The target video is eligible when it falls beyond a preregistered baseline band or the creator acted because they believed it had failed. Those two groups may differ and should be reported separately. The second captures real decision behavior even when the statistical label is weak.

Record when the judgment occurred. A creator who acted after one hour had a different evidence set from one who waited seven days. The timing belongs in the analysis, not in a moral ranking of patience.

Four evidence layers, none treated as ground truth

LayerWhat it contributesWhat it cannot settle
Creator accountIntent, audience knowledge, business constraints, perceived cause, emotional urgency, evidence used, and first action.A self-report cannot recover every sequence perfectly and may change after the outcome is known.
Native analyticsExposure, chose-to-view or equivalent fields, watch behavior, retention where available, engagement, account status, and timing.A dip or low reach does not name one edit decision as the cause.
Blinded human reviewIndependent observations of the opening, context, captions, continuity, pacing, payoff, and reasonable alternative explanations.Reviewer agreement is not audience behavior, and reviewers may miss creator or cultural context.
ViralJury analysisA repeatable product output that can be compared with humans and the creator under a fixed version.The product cannot serve as its own gold standard or reveal a platform's recommendation logic.
External contextEligibility notices, publication errors, topic demand, seasonality, policy changes, and concurrent account events.The study will never observe every external variable.

Action bias is a hypothesis, not a diagnosis

Action bias describes situations in which acting can feel preferable to waiting even when the evidence for action is weak. A well-known study found this pattern among elite football goalkeepers facing penalties. That result motivates a question about creator decisions under uncertainty; it does not prove creators behave the same way.

The study measures the proposed mechanism directly. Creators report urgency, uncertainty, perceived control, and the need to act using short, plain-language items developed for this context. It should not borrow a clinical burnout inventory or decision scale casually, relabel the score, and imply psychological diagnosis. Any validated instrument requires appropriate permission, administration, and interpretation.

Call an action a panic response only when the preregistered conditions are met: high reported urgency, a broad or irreversible first action, and little evidence recorded for that action at the time. Publish the component fields so readers can challenge the label.

The alignment framework

Alignment compares the action with the available evidence. It does not rank the creator's worth or assume the evidence is complete.

CategoryRuleExample
Directly alignedThe first action targets the primary, high-confidence issue supported by more than one evidence layer.A clear early retention drop and blinded missing-context finding lead to a specific opening-context revision.
Reasonably alignedThe action addresses a supported secondary issue or uses a broader but defensible fix.Caption/speech mismatch is observed and the creator retimes the opening caption system.
Partly alignedThe creator identifies the right area but the change does not isolate or fully address the evidence.A payoff problem leads to a shorter video without moving or clarifying the payoff.
Different targetThe action changes a variable not connected to the registered evidence, without implying the creator was irrational.The observed issue is a broken promise, while the first action changes hashtags.
Potentially costlyThe action removes evidence, adds policy or business risk, or makes a sweeping change unsupported by the record.The creator deletes all prior videos after one weak result.
IndeterminateEvidence conflicts, confidence is low, external factors are plausible, or creator context changes the interpretation.The video is structurally coherent but recommendation eligibility was unclear at the time.
No actionThe creator waits, collects more data, or makes no first change.This is analysed as a real response, not treated as missing data.

The survey captures sequence before hindsight rewrites it

Recruit soon after the target event and ask the perceived cause before showing any structural categories or analysis. Then capture the first action, evidence source, timing, and expected metric. Only after those responses are locked should ViralJury or a reviewer report be shown.

A follow-up records what happened to the target and next comparable video, whether the creator still endorses the action, and what they would do differently. These outcomes describe learning and hindsight. They do not validate the original action because performance naturally varies and the next upload is not a controlled counterfactual.

Open-text responses should be paraphrased or quoted only with explicit permission. Themes such as platform distrust, metric fixation, experimentation, or exhaustion need a frozen codebook, independent coders, and language that does not ridicule the participant.

Study sequence

Screen consent, age, media rights, analytics availability, comparison set, target-event timing, and elevated-risk content.

Lock the creator's baseline, flop definition, initial diagnosis, evidence used, urgency items, and first action before any review is revealed.

Collect the native analytics export or redacted screenshots with metric definitions, timestamps, and account-status context.

Run blinded independent human review and a version-locked ViralJury analysis without access to the creator survey or public result.

Adjudicate evidence conflicts under the frozen alignment framework while preserving every original rating and the indeterminate option.

Collect a preregistered follow-up, publish aggregate results and uncertainty, then destroy restricted media on the declared schedule.

Analysis keeps urgency, evidence, and action separate

The primary report shows distributions: perceived causes, evidence sources, judgment delays, first actions, alignment categories, and indeterminate cases. Cross-tabulations can compare whether actions based on native analytics, comments, intuition, peers, platform documentation, or automated review differ in alignment, but those associations remain subject to selection.

A preregistered multilevel model may examine whether urgency is associated with broader action after accounting for experience band, creator-level clustering, performance gap, platform, and format. The outcome should remain transparent; an opaque composite 'panic score' would recreate the black-box problem the study claims to examine.

Follow-up performance is secondary. Regression to the mean is especially important because the target entered after an extreme result. A better next video cannot be credited to the creator's action without a stronger counterfactual.

Ethical and publication gates

  • Participation, product use, and permission to publish identifiable case material are separate choices.
  • Creators can withdraw under a process stated before collection, including what happens after de-identification or publication.
  • Raw videos, analytics screenshots, handles, and identity keys are access-controlled and deleted on a declared schedule.
  • Subgroups too small for safe reporting are merged or suppressed, and regional comparisons are planned rather than added after an interesting pattern appears.
  • Low agreement, conflicting evidence, and indeterminate cases are published, not cleaned away.
  • The report critiques decisions and evidence, never the creator's intelligence, character, or mental health.
  • ViralJury results are evaluated against independent evidence and are not used as the gold standard for their own validation.

Sources and protocol foundations

The action-bias hypothesis is motivated by Bar-Eli and colleagues' study of goalkeeper decisions. The study context is different, so the creator hypothesis must be tested rather than assumed.

The follow-up analysis uses the same regression-to-the-mean warning described in Assessing regression to the mean effects in health care initiatives: an extreme baseline can create apparent improvement without an effective intervention.

YouTube's current guidance says performance reflects appeal, engagement, and satisfaction signals and tells creators to use their own analytics: Understand your content performance for recommendations. Its separate recommendation overview describes personalized signals rather than one creator-side formula.

Consent and withdrawal procedures should be specified before collection, following applicable review and resources such as HHS withdrawal guidance. The confirmatory plan belongs in a timestamped repository such as OSF Registrations.

Frequently asked questions

Does ViralJury decide what a video actually needed?

No. The analyzer is one versioned evidence layer. Native analytics, blinded human review, creator intent, and external context may agree or conflict. The protocol preserves those conflicts and allows an indeterminate result.

What counts as a panic response?

Only the preregistered combination of high self-reported urgency, a broad or irreversible first action, and little recorded evidence for that action. The label is descriptive within the study and is not a psychological diagnosis.

Can the study prove a creator chose the wrong fix?

Usually not in a causal sense. It can show whether the first action addressed the evidence available under the frozen framework. Unobserved audience, business, or platform context may still make a different action reasonable.

Has ViralJury measured a creator diagnosis gap?

No. No participants, videos, alignment ratings, or follow-up outcomes exist from this study yet. Any percentage would be a placeholder, not a finding.