Measured copy patterns · public sample

Diabetes VSL Copywriting Dataset

A structured view of the hooks, mechanisms, pain frames, promises, proof moves and CTAs found across processed diabetes VSL and ad transcripts.

3,350

high-confidence patterns

11

processed transcripts

11

distinct products

11 / 0

VSL / ad sources

What this dataset measures

The unit is a derived copy pattern, not a whole sales script. Each row labels what job the pattern performs, where it appears in the creative, which emotional tones it uses and how confident the extraction pass was. That makes the sample useful for structure analysis, taxonomy design, creative research and hypothesis generation without turning it into a verbatim swipe file.

Preview the actual fields

TypePositionCanonical patternConfidence
authorityproblem agitationSpeaker has spoken at international medical congresses, contributed to scientific journals, and consulted for global health initiatives0.90
avatarproblem agitationSomeone suffering from diabetes who wants to eat favorite foods, share family meals, and live without blood sugar anxiety0.84
ctaproblem agitationClick the green button now to secure the package before time runs out0.97
hookproblem agitationUniversal exposure: nearly everyone who eats common foods has already been affected by metabolic biofilm0.89
mechanismproblem agitationDesert-grown extract formulated to dissolve metabolic biofilm and restore pancreatic function0.95
painopeningBlood sugar spiraling out of control despite healthy eating and medication adherence0.98
promiseproblem agitationComprehensive support to eliminate metabolic biofilm, repair pancreas, and maintain optimal blood sugar levels0.95
social_proofopeningResearchers from Imperial College London and University of Columbia Cairo have confirmed the metabolic biofilm discovery0.88
tacticproblem agitationContrast between forbidden foods in Western medicine and thriving elders consuming those same foods0.89
urgencyproblem agitationTime-limited stock with color-coded button that turns red when supply is exhausted0.95

How rows qualify

A source transcript must be active, English, reviewed, fully chunked and fully embedded. Its extraction pass must contain at least eight units, with at least 75% scoring 0.80 or higher. The export then rejects near-verbatim source language and deduplicates matching canonical patterns inside each niche and pattern type.

Counts describe observed corpus structure, not proof that a phrase causes conversion, and they do not validate medical or commercial claims made inside an advertisement.

Need the full cross-niche edition?

The commercial release combines every mature niche in one consistent CSV/JSONL schema, includes checksums and a data dictionary, and remains de-identified.

View the full dataset