The full methodology
The transparent framework behind every TV Intelligentsia score.
Download PDF (v1.3.1)Why this document exists #
1.1 The measurement gap in entertainment
The entertainment industry has extensive infrastructure for measuring who watches (Nielsen), what is in demand (Parrot Analytics), and what people say they liked (Rotten Tomatoes, IMDb). It has far less transparent infrastructure for assessing the cognitive demands, educational affordances, and craft properties present in the work itself. TV Intelligentsia was built to fill that gap. TVI is the credibility layer for what to watch.
1.2 The credibility crisis in existing ratings
Every major rating system in entertainment has a documented structural problem.
Rotten Tomatoes binarizes critic opinion into a Fresh/Rotten vote, then aggregates the votes into a percentage. The methodology rewards consensus and obscures intensity. A film that nine of ten critics mildly approve of scores higher than one that four of ten critics consider a masterpiece.
Metacritic weights critic scores through an undisclosed proprietary formula, which by definition cannot be evaluated, debated, or reproduced.
IMDb and Letterboxd rely on user-generated ratings vulnerable to review bombing, brigade voting, and cultural-warfare campaigns. The pattern has been visible on major franchise releases from 2019 forward.
Common Sense Media focuses on age-appropriateness, a real service, but a different construct than content quality or cognitive value.
Nielsen measures attention volume, not attention quality.
TVI is designed as a published, version-stamped, change-logged alternative with named human accountability. The comparison here describes TVI's design choice; it is not a claim that every other service lacks all public methodology material.
1.3 What this document is and is not
This is a published methodology, version-stamped, change-logged, and scheduled for quarterly review. It is a transparent account of the current scoring standard, grounded in published cognitive science and designed to be reproducible by qualified reviewers. Every published C/E/Q breakdown must reproduce its displayed composite. As of August 2, 2026, every current catalog row meets that integrity rule; any future mismatch is withheld from dimension-level publication until human correction.
It is not a claim of peer-reviewed validation. That is future work. TVI's current proof is a public rubric, versioned calculations, visible reasoning, named accountability, and a corrections process. Inter-rater reliability and any external validation remain uncompleted future phases and are not current claims.
1.4 The founder's measurement credential
Jordan Robinson holds an MD and a Master of Public Health and authored the TVI Score framework. His public-health research training informs the framework's versioning, explicit definitions, calculation rules, and stated limitations. Those credentials do not constitute external validation of TVI's instrument. The methodology stands or falls on its published logic, disclosed evidence, calibration record, and correction process.
The IQ Score framework #
2.1 The IQ Score defined
The TVI IQ Score is a reviewed content rating on a 0 to 200 scale using three weighted dimensions: Cognitive Stimulation, Educational Value, and Craft & Quality. It describes properties of the work: its cognitive demands, the knowledge and skills it makes available, and the quality of its execution. It does not claim what viewing caused in a particular person.
The IQ Score is a content rating. It is not a measurement of viewer intelligence. It is not a safety rating. It is not an age-appropriateness rating.
2.2 The formula
Note on naming. The third dimension was named Entertainment Quality through v1.2 and carried the abbreviation EQ. As of v1.3 it is named Craft & Quality (CQ). The name change describes the dimension more accurately, since it measures craft in service of the work rather than entertainment value in isolation, and it retires the collision between the old EQ abbreviation and the separately developed Emotional Intelligence dimension in Section 5, which uses the display label "EI Score". The weight (25%) and the formula are unchanged.
Dimension scores are multiplied by their weights, summed, multiplied by four to produce a 0 to 200 scale, and rounded to the nearest integer. For nonnegative values, exact halves round upward. The same rule applies when the four Cognitive Stimulation or Craft & Quality sub-metrics are averaged. Maximum composite: 200.
Score-integrity note, August 2, 2026. A three-dimension breakdown is published only when its values reproduce the displayed composite under this formula. Every current catalog row now meets that standard. If a future row fails it, dimension detail is withheld and human correction is required; no replacement judgment is generated automatically.
2.3 Why these weights
Cognitive Stimulation receives the highest weight (40%) as a published design judgment. TVI gives primary emphasis to observable demands for tracking, inference, working memory, and conceptual integration. The cited literature informs those constructs, but does not validate this weight or establish a viewer outcome. Four sources inform the construct directly:
- Cognitive Load Theory (Sweller, 1988; Sweller, van Merriënboer, & Paas, 1998) establishes that the mental architecture activated by content varies dramatically with structural complexity, and that the load a viewer must process determines the cognitive work performed.
- Narrative Transportation Theory (Green & Brock, 2000) informs the distinction between structural features of a narrative and a viewer's reported engagement. TVI does not treat it as proof of retention from a particular title.
- The Limited Capacity Model of Motivated Mediated Message Processing (Lang, 2000; Lang, 2006) links media structural features directly to cognitive resource allocation.
- Madigan et al. (2020) reviews associations among the quantity, content, and context of screen use and child language. TVI cites it as background only, not as validation of a title score, a weight, or a predicted developmental outcome.
Educational Value receives the second-highest weight (35%) because accurate knowledge, visible reasoning, emotional-process modeling, usable behavior, and transfer support are central to TVI's assessment of value. The literature establishes that media can support comprehension and learning under some conditions; it does not validate this particular weight or prove that a specific viewer learned, retained, or benefited. Fisch (2004), Desmond & Dillman Carpentier (2019), and Butler, Zaromb, Lyle, & Roediger (2009) inform the construct while leaving those outcome questions open.
Craft & Quality receives the lowest weight (25%) as an editorial design choice. TVI gives more composite weight to cognitive demand and educational material while retaining a substantial role for execution. Excluding craft entirely would allow technically weak work to outrank accomplished work solely because of subject matter. The 25% weight rewards craft without letting it dominate.
2.4 The weighting rationale as a design decision
The weights represent informed judgment, not empirical optimization. They prioritize observable work properties in this order: cognitive demand, educational material, then entertainment craft. The cited literature does not dictate the percentages and the dimensions do not predict an effect on an individual viewer. TVI publishes the weights so they can be evaluated, debated, and adjusted in future versions.
Dimension definitions and sub-metrics #
Each dimension is scored 0 to 50. Cognitive Stimulation and Craft & Quality are each the average of four sub-metrics, each scored 0 to 50. Educational Value is the sum of five sub-metrics, each scored 0 to 10, one for each of the five forms of learning the dimension measures. Thirteen sub-metrics across three dimensions constitute the full scoring rubric.
3.1 Cognitive Stimulation, 40% weight
How hard the viewer's brain works during viewing. Measures the cognitive resources required to follow, process, and engage with the content.
| Sub-metric | What it measures | High-score exemplars | Grounding |
|---|---|---|---|
| Narrative Complexity | Number and interconnection of plot threads, temporal structures, character arcs, thematic layers | Interweaving storylines, non-linear timelines, unreliable narrators, thematic recursion. The Wire, Dark | Mittell (2015), Complex TV |
| Dialogue Density | Lexical sophistication, information load per line, subtext, conversational complexity | Dialogue requiring active parsing, domain vocabulary, double meanings. The West Wing, Succession | Lang (2000), LC4MP |
| Cognitive Load | Degree to which the viewer must actively construct meaning rather than passively receive it | Shows where critical information is lost if you look at your phone. Severance, Westworld | Sweller (1988), Cognitive Load Theory |
| Conceptual Novelty | Introduction of ideas, frameworks, or perspectives the viewer is unlikely to have encountered | Content that introduces a substantially different conceptual frame for its subject. Black Mirror, Cosmos | Berlyne (1960), curiosity as a cognitive driver |
3.2 Educational Value, 35% weight
What the work makes available. Educational Value describes accurate information, visible reasoning, emotional-process modeling, life-skill depiction, and transfer architecture present in the work. It is not limited to academic subject matter. The five sub-metrics rate those observable properties. They do not establish that a particular viewer learned, retained, or changed.
Each sub-metric below is scored 0 to 10. The five sum to the dimension's 0 to 50 score.
| Sub-metric | What it measures | High-score exemplars | Grounding |
|---|---|---|---|
| Academic Content | Volume and accuracy of verifiable real-world information conveyed: facts, vocabulary, history, science, with fidelity to established fact and science | Cosmos, Chernobyl, Blue Planet | Fisch (2004); Mutz & Goldman (2010) |
| Emotional Intelligence | How precisely the content depicts empathy, self-awareness, emotional regulation, and the naming of feelings | Inside Out, Bluey, BoJack Horseman | Immordino-Yang & Damasio (2007) |
| Critical Thinking | Whether the content exercises cause-and-effect reasoning, inference, problem-solving, and moral reasoning rather than delivering conclusions | The Wire, Mr. Robot, Avatar: The Last Airbender | Halpern (1998) |
| Life Skills | Whether the content demonstrates behaviors, decisions, or frameworks a viewer can apply to their own life and work | Ted Lasso, The Good Place, Daniel Tiger's Neighborhood | Bandura (1986), social cognitive theory |
| Knowledge Transfer | Observable support for retrieving, applying, integrating, or generalizing an idea across contexts | House M.D., Band of Brothers, Cosmos | Butler et al. (2009) |
3.3 Craft & Quality, 25% weight
Craft and engagement. Measures the technical and artistic quality of the content as a piece of made entertainment.
| Sub-metric | What it measures | High-score exemplars | Grounding |
|---|---|---|---|
| Emotional Range | Breadth and depth of emotional experience represented and invited by the work | Work that moves beyond a single emotional register. Six Feet Under, BoJack Horseman | Nabi & Green (2015) |
| Narrative Arc Completion | Degree to which the story delivers on its structural promises | Satisfying resolution of character arcs, thematic payoff. Breaking Bad vs mid-season algorithmic filler | Brewer & Lichtenstein (1982) |
| Production Value | Cinematography, sound design, editing, visual effects, score | Production choices that serve the story rather than substitute for it. Shogun, Chernobyl | Industry-standard craft assessment; expert judgment |
| Audience Retention | Degree to which the content sustains engaged attention across its runtime | Content that earns the viewer's time minute-by-minute, not just episode-by-episode | Green & Brock (2000) |
The SEL dimension (children's content) #
4.1 SEL defined and why it is separate
Social-Emotional Learning is scored 0 to 50 and reported alongside the IQ Score as a distinct measure. It does not feed into the IQ Score formula.
SEL measures developmental appropriateness and emotional-skill modeling, a different construct than cognitive or educational quality. A children's show can score 160 IQ and 8 SEL (highly stimulating, minimal social-emotional modeling) or 88 IQ and 48 SEL (minimal cognitive demand but exceptional emotional-development content). Conflating the two would obscure the signal each provides.
4.2 The CASEL framework
Under the current standard, TVI's SEL dimension uses the CASEL (Collaborative for Academic, Social, and Emotional Learning) framework, the same framework adopted by K-12 schools nationwide. The five competencies are scored and preserved for current-standard kids scores. Legacy SEL values are historical total-level judgments only; competency-level review remains pending.
| CASEL competency | What it measures in content |
|---|---|
| Self-Awareness | Does the content model characters recognizing their own emotions, strengths, and limitations? |
| Self-Management | Does the content model emotional regulation, impulse control, and goal-directed behavior? |
| Social Awareness | Does the content model empathy, perspective-taking, and appreciation for diversity? |
| Relationship Skills | Does the content model healthy communication, cooperation, and conflict resolution? |
| Responsible Decision-Making | Does the content model ethical reasoning, consequential thinking, and constructive choices? |
4.3 The credentialed reviewer
The TVI Kids methodology is shaped and reviewed by Cordelia Witty, EdS., NCSP, an Education Specialist and Nationally Certified School Psychologist. Her credential anchors methodology review. It does not mean she reviewed every catalog title. Her name appears as a page-level reviewer only where she personally reviewed and approved the work. The methodology does not constitute an individualized assessment or establish that viewing produces a particular developmental outcome.
4.4 Anchor calibration scores
SEL scoring is anchored to reference titles that establish the scale. Anchor scores are pulled live from the TVI database at tvintelligentsia.com/explore at time of publication and may evolve with rescoring; the live database is always authoritative.
- Daniel Tiger's Neighborhood, SEL 48. High anchor. Explicit emotional-literacy instruction with clear SEL scaffolding by design.
- Bluey, SEL 46. Exceptional modeling of emotional regulation, family dynamics, imaginative play, and age-appropriate problem-solving.
- Sesame Street, SEL 44. Broad SEL range with strong modeling across all five CASEL competencies.
- Ms. Rachel (Songs for Littles), SEL 44. High. This legacy total reflects strong language and emotional-vocabulary modeling for early-toddler audiences; competency-level SEL review remains pending.
- Teletubbies, SEL 10. Low-mid range. Visual and tonal warmth without direct emotional-skill modeling.
- Cocomelon, SEL 8. Low anchor. Minimal emotional-regulation modeling, low social-interaction depth, high stimulation rate without SEL resolution.
- Baby Shark, SEL 6. Low anchor. Sensory-driven content with minimal character interaction to model social-emotional skills.
4.5 The SEL calibration flag
To prevent Educational Value from being assessed too narrowly, for instance, counting only academic content and ignoring emotional-intelligence or life-skills transfer, every scoring session runs an automated flag: any children's title where SEL is at least 40 and EV is no more than 25 is surfaced for manual review. A discrepancy of this size almost always indicates that educational value was scored against an academic-only ceiling rather than the five-dimensional rubric (academic content, emotional intelligence, critical thinking, life skills, knowledge transfer). Flagged titles are reviewed by the scoring team, not auto-corrected.
EI dimension research archive #
This section preserves a research concept first documented in v1.1. The EI dimension is not an operational TVI product, is not scored in the public catalog, and has no current signatory or advisory-board authority. Any future launch would require a new versioned methodology, named current accountability, evidence review, and public disclosure.
5.1 What EI measures
The proposed TVI EI dimension would describe the relational and emotional sophistication present in a work: how precisely it portrays empathy, self-awareness, regulation, and human relationships. It would rate observable features of the work, not claim to develop emotional intelligence in a viewer.
Note on naming. Through v1.2 the IQ Score formula used the abbreviation EQ for the third dimension, then named Entertainment Quality. As of v1.3 that dimension is Craft & Quality (CQ), so the collision is retired. The Emotional Intelligence dimension continues to use the display label EI Score on all public-facing surfaces. In version history prior to v1.3 it is referred to as the EQ dimension for continuity with v0.1 planning references.
5.2 Why EI is a separate dimension
The IQ Score's Cognitive Stimulation and Educational Value dimensions capture intellectual complexity and transfer support. Craft & Quality captures craft and engagement. None of these fully isolates how a title models or calls on empathy, self-awareness, emotional regulation, and relational understanding.
A show can score 185 on the IQ Score while presenting characters with no emotional interior life, procedural excellence without emotional depth. A show can score 112 on the IQ Score while featuring some of the most emotionally intelligent character writing in the medium. The EI Score captures the latter signal independently of the former.
5.3 Theoretical grounding
The proposed dimension draws construct vocabulary from psychology and narrative research. Those sources do not validate the proposed TVI instrument or establish that watching a specific title changes emotional competence.
Salovey and Mayer's four-branch model (1990, 1997) defines emotional intelligence across four hierarchical domains: perceiving emotions (recognizing emotional signals in faces, voices, and content), using emotions (harnessing emotions to facilitate thought), understanding emotions (comprehension of emotional complexity and transitions), and managing emotions (regulation of emotional experience). TVI's EI Score evaluates how effectively a title's characters model and navigate each of these domains.
Goleman's emotional intelligence framework (1995, 2006) maps emotional intelligence onto five practical competencies, self-awareness, self-regulation, motivation, empathy, and social skills, that parallel and extend the CASEL framework used in TVI's SEL dimension. Where SEL captures developmental modeling for children, EI captures adult-grade emotional intelligence in all content.
Mar and Oatley's simulation hypothesis (Mar et al., 2006; Mar, 2011) proposes that fiction can function as a simulation of social worlds. TVI uses this as background for identifying perspective-taking in a work, not as proof that a TVI score predicts a viewer outcome.
Kidd and Castano (2013) reported short-term Theory of Mind results in experiments involving literary fiction. TVI does not extrapolate that result to television viewers or use it as evidence that a proposed EI score would produce or predict a personal effect.
Zillmann's affective disposition theory (1994, 2000) informs analysis of how a work structures alignment with characters. It does not validate a TVI outcome claim.
5.4 The five sub-dimensions
The EI Score is scored 0 to 50, as the average of five sub-dimensions each scored 0 to 50. Unlike the IQ Score's three main dimensions, the EI Score's five sub-dimensions are weighted equally.
Sub-dimension 1, Empathy Modeling
Does the content portray characters who demonstrate the capacity to recognize, understand, and share the feelings of others? This sub-dimension evaluates whether empathy is modeled as a functional, practiced capacity, not merely referenced as a value. High-scoring content shows characters actively taking the perspective of others in conflict and adjusting their responses based on that understanding.
High-anchor examples: Inside Out, BoJack Horseman (later seasons), Six Feet Under.
Low-anchor examples: Content in which characters' emotional states are functional plot devices rather than developed interior lives.
Sub-dimension 2, Self-Awareness Depiction
Does the content portray characters with accurate, developing self-knowledge? This sub-dimension evaluates whether characters demonstrate awareness of their own emotional patterns, blind spots, defense mechanisms, and growth over time. Critically, self-awareness depiction distinguishes between characters who know themselves and characters who merely narrate themselves, the latter being a common substitution.
High-anchor examples: The Sopranos (Tony's therapy arc), Fleabag, Better Call Saul.
Low-anchor examples: Content in which characters' stated self-awareness is contradicted by behavior without the contradiction being recognized narratively.
Sub-dimension 3, Emotional Regulation in Characters
Does the content portray characters who demonstrate the capacity to manage emotional states, including the realistic portrayal of failure to regulate? This sub-dimension does not reward stoicism or emotional suppression. It rewards authentic portrayal of emotional regulation as a practiced, imperfect, developmental capacity. High-scoring content shows characters using named strategies (explicitly or implicitly), experiencing the costs of failed regulation, and developing regulatory capacity across a narrative arc.
High-anchor examples: Bluey (parenting regulation arcs), Succession (dysregulation as character study), Mare of Easttown.
Low-anchor examples: Content in which emotional dysregulation is presented as a character trait with no consequence or arc, or in which regulation is presented as unproblematic stoicism.
Sub-dimension 4, Relational Complexity
Does the content portray relationships with genuine psychological depth? This sub-dimension evaluates whether the interpersonal relationships in the content demonstrate the actual complexity of human connection, including ambivalence, unresolved conflict, repair, rupture, implicit communication, and the layered history that characterizes real relationships. High-scoring content portrays relationships that are neither idealized nor pathologized, but recognizably, messily human.
High-anchor examples: The Americans, Parenthood, Marriage Story.
Low-anchor examples: Content in which relationships serve narrative functions without psychological interiority, or in which relational complexity is flattened to archetypal roles.
Sub-dimension 5, Emotional Vocabulary and Specificity
Does the content use specific, differentiated emotional language in dialogue, narration, or visual portrayal? This sub-dimension evaluates whether the content distinguishes between related emotional states rather than operating at the level of generic emotional reference. High-scoring content treats emotional language as meaningful and precise. The proposed score would describe that property, not predict a viewer outcome.
High-anchor examples: Inside Out (the Joy/Sadness distinction as the film's central argument), Normal People, Afterlife (Ricky Gervais).
Low-anchor examples: Content in which emotional states are named only at the level of happy/sad/angry/scared.
5.5 The EI anchor calibration
The following titles are archived proposals, not public scores or approved calibration anchors. They cannot be used in a product, ranking, or claim without a future versioned review.
Archived proposed high anchors, not approved scores:
- Inside Out (Pixar), EI 49/50 (archived proposal). The film's central narrative argument represents multiple emotional states, the function of sadness, and emotional integration. This is an observation about the work, not a clinical claim or viewer outcome.
- BoJack Horseman, EI 47/50 (proposed). Perhaps the most psychologically sophisticated depiction of defense mechanisms, avoidance, and self-sabotage in the animated medium.
- The Sopranos, EI 45/50 (proposed). The therapy-session structure provides a systematic container for self-awareness depiction across the series run.
Archived proposed low anchors, not approved scores:
- High-stimulation procedural with flat character emotional interiors, EI 12/50 (archived proposal only).
- Reality competition format with emotional portrayal reduced to strategic performance, EI 8/50 (archived proposal only).
5.6 Current governance status
There is no current EI signatory, advisory board, or credentialed panel. The material in this section is a transparent archive of a proposed construct, not evidence of current clinical or editorial authority. If TVI resumes this work, the new methodology version must name the actual accountable people and the evidence they personally reviewed.
5.7 Launch criteria
The EI dimension may launch publicly only after a new methodology version defines the construct boundary, names current human accountability, documents the evidence review, publishes calibrated anchors, and clearly separates work properties from viewer outcomes. No launch is scheduled.
Scoring protocol #
6.1 Unit of analysis
TVI's unit-of-analysis policy treats a non-anthology series as a complete body of work across all seasons, rather than episode by episode. Rationale: episode-level scoring introduces volatility that obscures aggregate signal, and the consumer question is should I watch this show, not should I watch this episode.
The live catalog is still in migration: legacy season-range rows remain for some non-anthology shows. Readers may therefore encounter both a whole-series row and one or more season-range rows until the separately governed consolidation is approved and executed.
Anthology exception. Anthology series where each season is a different narrative work, True Detective, Fargo, Black Mirror, The White Lotus, retain season-level entries, because each season is scored as a distinct work. Outside those anthologies, one canonical full-run entry is the target state, not yet a universal description of the current catalog.
6.2 Scoring process
- Evidence standard. Under the current standard, in force since August 2026, every score is grounded in direct engagement with the work. Scores published under our earlier workflow, before the current standard took effect in August 2026, carry the legacy label; see the migration note on our corrections page. Flagship reviews, Essential designations, and premiere-session scores are based on full viewings, and no experiential claim is published beyond what has actually been seen.
- Independent sub-metric scoring. Under the current standard the named rater scores each of the thirteen sub-metrics using the rubric definitions in Section 3, and the thirteen judgments are preserved as the score's receipt; legacy scores do not carry a verifiable current-standard receipt. Cognitive Stimulation and Craft & Quality sub-metrics are scored 0 to 50; Educational Value sub-metrics are scored 0 to 10.
- Dimension aggregation. Cognitive Stimulation and Craft & Quality are each the mean of their four sub-metric scores. Educational Value is the sum of its five sub-metric scores.
- Formula application. The IQ Score is computed via the formula in Section 2, and every current catalog row reproduces its composite. A future row that does not cannot publish dimension detail until human review; no formula-derived replacement judgment is created automatically.
- SEL scoring (children's content only). Under the current standard, kids titles carry a separate SEL Score produced from the five CASEL competencies in Section 4, and those competency judgments are preserved. Earlier-workflow kids titles display a legacy SEL total reviewed only at the total level; competency-level review remains pending. Cordelia Witty shapes and reviews the methodology and personally reviews only the pages and TVI Kids Essential designations attributed to her.
- EI scoring. Not operational. The archived research concept in Section 5 is not displayed or scored publicly.
- Internal consistency review. Scores are checked against internal consistency benchmarks, for example, a title cannot score 45/50 on Cognitive Stimulation overall while scoring 10/50 on Narrative Complexity without flagging for a sub-metric audit.
- Calibration sweep. All kids titles are run against the SEL calibration flag described in Section 4.5.
6.3 Reviewer qualifications
TVI does not claim a current credentialed panel. Jordan Robinson, MD, MPH, is the accountable editor for catalog scores and the author of the TVI Score framework. Cordelia Witty, EdS., NCSP, shapes and reviews the TVI Kids methodology; her name appears as a page-level reviewer only where she personally reviewed the work. Any future independent reviewer must have explicit qualifications, agree to the published rubric, and be identified on the work they actually reviewed.
6.4 Inter-rater reliability (Phase 2)
TVI acknowledges that a single accountable catalog editor limits claims about reproducibility. The Phase 2 reliability program uses independent reviewers, the same frozen scored object and instrument, ICC(A,1) for absolute agreement, numerical-difference and Bland-Altman reads, and tier agreement. The initial eight-title run is instrument debugging; a pre-specified adult pilot of at least 60 double-scored works is required before any public reproducibility claim is considered. Reliability evidence will not be described as construct validation or as proof of viewer outcomes.
6.5 Score integrity protocol
Scores published on tvintelligentsia.com are authoritative. Planning documents, marketing materials, internal drafts, and external citations are not publishable sources for scores. When TVI content cites a score, the score must be verified against the live database at the time of publication. The database is the record; anything else is a reference.
Score categories and database distribution #
7.1 The five tiers
| Range | Tier | Definition |
|---|---|---|
| 160-200 | Masterclass | Exceptional cognitive and educational substance, delivered with craft that sustains close attention |
| 130-159 | Stimulating | Significantly challenges the viewer intellectually |
| 100-129 | Competent | Meaningful engagement above passive consumption |
| 70-99 | Passive | Minimal cognitive demand, entertainment-driven |
| 0-69 | Numbing | Negligible intellectual engagement |
7.2 Current database distribution
As of August 23, 2026, the TVI database contains 2,606 scored titles (2,379 adult, 227 children's). Distribution across tiers:
| Tier | Count | Share |
|---|---|---|
| Masterclass (160+) | 681 | 26.1% |
| Stimulating (130-159) | 896 | 34.4% |
| Competent (100-129) | 868 | 33.3% |
| Passive (70-99) | 142 | 5.5% |
| Numbing (<70) | 18 | 0.7% |
7.3 Distribution analysis
The database skews toward higher-quality content because the initial corpus was built from titles with demonstrated cultural significance, critical acclaim, or sustained audience interest. As the database expands to include more platform-filler content, procedural network television, algorithmic filler, unscripted programming that currently sits outside the corpus, the distribution will shift downward. This is expected and methodologically appropriate: TVI rates what exists, not what it wishes existed.
Supplementary dimensions (banked) #
One additional dimension is designed and partially implemented. It is documented here to signal methodological depth and roadmap intent, and to prevent future-version additions from appearing ad hoc. (The previously-banked EQ/Emotional Intelligence dimension was promoted to its own Section 5 in v1.1.)
8.1 Cinematic Score, Music and Soundtrack
Scored 0 to 50. Measures composition quality, emotional impact, thematic integration, and memorability of a title's score and soundtrack. Currently populated for 428 titles where soundtrack is a significant artistic element. Reported alongside the IQ Score as a distinct dimension on qualifying titles. Does not feed into the IQ Score formula. Does not apply universally; titles without meaningful soundtrack presence do not carry a Cinematic Score.
What TVI does not measure #
Defining the boundary of the methodology is as important as defining what is inside it.
TVI does not measure age-appropriateness. That is Common Sense Media's domain and they do it well. TVI measures quality and cognitive value, which is a different construct. A show can be age-appropriate and cognitively numbing, or age-inappropriate and intellectually masterful.
TVI does not measure popularity or audience sentiment. That is Nielsen's, IMDb's, and Rotten Tomatoes' domain. A show can be wildly popular and score 82 or relatively obscure and score 189. Popularity and cognitive value are independent axes.
TVI does not measure viewer enjoyment. A viewer may genuinely enjoy a Passive-tier show more than a Masterclass-tier show on a given evening. The IQ Score rates the content, not the experience. Enjoyment is a separate question that TVI explicitly does not answer.
TVI does not make clinical claims about individual viewers. The IQ Score does not predict that watching a 200-rated show will make a specific viewer smarter. It measures the cognitive, educational, and entertainment value of the content itself. Individual cognitive effects depend on attention, context, co-viewing, prior knowledge, and other variables outside TVI's scope.
TVI does not measure political ideology, representation, or values alignment. These are legitimate analytical frames. They are not the frame TVI measures.
Dispute process #
A published methodology is only credible if it can be challenged.
10.1 Score disputes
Any member of the public, any creator, and any platform may dispute a specific score by submitting a written challenge to methodology@tvintelligentsia.com. A dispute must identify the title, the current score, the sub-metric(s) the challenger believes are misscored, and a reasoned argument grounded in the rubric.
A dispute is not a request to raise or lower a score on taste grounds. A dispute is a claim that the rubric was misapplied. Disputes grounded in taste, political disagreement, or marketing objection are acknowledged but not acted on.
10.2 Response timeline
TVI commits to acknowledging every dispute within fourteen days, and to responding with a reasoned decision within thirty days. Decisions fall into three categories:
- Rubric was misapplied. Score adjusted; change logged.
- Rubric was applied correctly, but the dispute surfaces a real edge case. Change-log entry flags the edge case for future rubric refinement.
- Dispute not supported. Score retained; response documents the rubric application.
10.3 Methodology challenges
Challenges to the methodology itself, sub-metric definitions, weights, scoring protocols, rubric logic, are welcomed and handled separately from score disputes. Substantive methodology challenges are incorporated into the quarterly review cycle and may produce versioned updates to this document.
10.4 Editorial independence
TVI accepts no payment, promotional consideration, or commercial incentive from studios, platforms, distributors, or talent agencies in exchange for scores. A score cannot be bought, negotiated, or withdrawn. This is stated in the founding operating principles and is not subject to dispute.
Methodology roadmap #
11.1 Near-term (Phase 1)
- Comparative cohort benchmarking, percentile scoring within genre, platform, decade, and age group.
- Attention-architecture mapping, scene length, cuts per minute, dialogue-to-silence ratios as additional cognitive-load signals.
- Published rubrics, this document.
- Preserve the proposed EI research concept as non-operational until a new evidence review and methodology version exists.
- Content-creator feedback loop, structured channel for creators to submit corrections to factual claims about their titles.
11.2 Medium-term (Phase 2)
- Any EI dimension work requires a new versioned methodology and named current accountability.
- Inter-rater reliability testing with qualified independent reviewers.
- Developmental-stage weighting for children's content.
- Scaffolding assessment, how content introduces complexity across a serialized run.
- Rewatch-decay modeling, how repeated viewing changes the cognitive signal.
- Genre-specific weighting experiments.
11.3 Long-term (Phase 3 and beyond)
- University-partnered EEG and fMRI studies correlating TVI scores with measured cognitive activity.
- Pre- and post-viewing knowledge delta testing for educational claims.
- Peer-reviewed publication of validation studies.
- Negative scoring, a cognitive-cost dimension distinct from the current baseline of zero.
Limitations and transparency #
Stating what TVI does not yet have is the section that builds the most credibility with sophisticated readers. It is stated here in full.
12.1 The database is 2,606 titles. This is sufficient for consumer-facing recommendations and genre-level aggregation. It is below the threshold for statistically robust claims about entire platform catalogs. Platform-level comparisons (for example, "Netflix's average IQ Score") are reported with explicit sampling caveats.
12.2 TVI does not have a credentialed reviewer panel. Jordan Robinson is the accountable editor for catalog scores. Cordelia Witty shapes and reviews the TVI Kids methodology and personally reviews only work attributed to her. Until inter-rater reliability is documented, TVI scores are human editorial ratings, not statistically validated measurements.
12.3 The weighting is judgment-based. The 40/35/25 split is informed by the literature cited in Section 2.3 but is not empirically optimized. Future versions may adjust weights through documented calibration and evidence review.
12.4 The score is a composite. Like any composite metric, it compresses multidimensional information into a single number. Per-dimension detail is published only where the stored values reproduce the displayed composite under the current formula. Every current row meets that rule; any future mismatch is held for human correction.
12.5 There is no peer-reviewed validation yet. TVI's phased credibility model is explicit: expert authority first (Phase 1), empirical evidence second (Phase 2), scientific validation third (Phase 3). This document is Phase 1 output. Claims of scientific validation in any TVI marketing or press material are unauthorized and should be reported.
12.6 The EI dimension is a non-operational research archive. No EI scores should appear on any TVI surface. The proposed framework and anchors are documented for transparency, not because they are approved, validated, or scheduled to launch.
12.7 The scoring corpus reflects reviewer access. Titles available on major streaming platforms and via physical or digital rental are accessible. Titles restricted to specific regional markets or behind one-off paywalls may be underrepresented.
References #
- Bandura, A. (1986). Social Foundations of Thought and Action: A Social Cognitive Theory. Prentice-Hall.
- Berlyne, D. E. (1960). Conflict, Arousal, and Curiosity. McGraw-Hill.
- Brewer, W. F., & Lichtenstein, E. H. (1982). Stories are to entertain: A structural-affect theory of stories. Journal of Pragmatics, 6(5-6), 473-486.
- Butler, A. C., Zaromb, F. M., Lyle, K. B., & Roediger, H. L. (2009). Using popular films to enhance classroom learning: The good, the bad, and the interesting. Psychological Science, 20(9), 1161-1168.
- CASEL (Collaborative for Academic, Social, and Emotional Learning). (2020). CASEL's SEL Framework: What Are the Core Competence Areas and Where Are They Promoted? Chicago, IL: CASEL.
- Desmond, R., & Dillman Carpentier, F. (2019). Media and the Well-Being of Children and Adolescents (updated edition). Oxford University Press.
- Fisch, S. M. (2004). Children's Learning from Educational Television: Sesame Street and Beyond. Lawrence Erlbaum Associates.
- Goleman, D. (1995). Emotional Intelligence: Why It Can Matter More Than IQ. Bantam Books.
- Goleman, D. (2006). Social Intelligence: The New Science of Human Relationships. Bantam Books.
- Green, M. C., & Brock, T. C. (2000). The role of transportation in the persuasiveness of public narratives. Journal of Personality and Social Psychology, 79(5), 701-721.
- Kidd, D. C., & Castano, E. (2013). Reading literary fiction improves theory of mind. Science, 342(6156), 377-380.
- Lang, A. (2000). The limited capacity model of mediated message processing. Journal of Communication, 50(1), 46-70.
- Lang, A. (2006). Using the limited capacity model of motivated mediated message processing to design effective cancer communication messages. Journal of Communication, 56(s1), S57-S80.
- Madigan, S., McArthur, B. A., Anhorn, C., Eirich, R., & Christakis, D. A. (2020). Associations between screen use and child language skills: A systematic review and meta-analysis. JAMA Pediatrics, 174(7), 665-675.
- Mar, R. A. (2011). The neural bases of social cognition and story comprehension. Annual Review of Psychology, 62, 103-134.
- Mar, R. A., Oatley, K., Hirsh, J., dela Paz, J., & Peterson, J. B. (2006). Bookworms versus nerds: Exposure to fiction versus non-fiction, divergent associations with social ability, and the simulation of fictional social worlds. Journal of Research in Personality, 40(5), 694-712.
- Mark, G. (2023). Attention Span: A Groundbreaking Way to Restore Balance, Happiness and Productivity. Hanover Square Press.
- Mayer, J. D., & Salovey, P. (1997). What is emotional intelligence? In P. Salovey & D. Sluyter (Eds.), Emotional Development and Emotional Intelligence (pp. 3-31). Basic Books.
- Mittell, J. (2015). Complex TV: The Poetics of Contemporary Television Storytelling. NYU Press.
- Mutz, D. C., & Goldman, S. K. (2010). Mass media. In J. F. Dovidio, M. Hewstone, P. Glick, & V. M. Esses (Eds.), The SAGE Handbook of Prejudice, Stereotyping and Discrimination (pp. 241-258). SAGE.
- Nabi, R. L., & Green, M. C. (2015). The role of a narrative's emotional flow in promoting persuasive outcomes. Media Psychology, 18(2), 137-162.
- Oatley, K. (2016). Fiction: Simulation of social worlds. Trends in Cognitive Sciences, 20(8), 618-628.
- Salovey, P., & Mayer, J. D. (1990). Emotional intelligence. Imagination, Cognition and Personality, 9(3), 185-211.
- Sweller, J. (1988). Cognitive load during problem solving: Effects on learning. Cognitive Science, 12(2), 257-285.
- Sweller, J., van Merriënboer, J. J. G., & Paas, F. G. W. C. (1998). Cognitive architecture and instructional design. Educational Psychology Review, 10(3), 251-296.
- Weissberg, R. P., Durlak, J. A., Domitrovich, C. E., & Gullotta, T. P. (2015). Social and emotional learning: Past, present, and future. In Handbook of Social and Emotional Learning (pp. 3-19). Guilford Press.
- Wolf, M. (2018). Reader, Come Home: The Reading Brain in a Digital World. HarperCollins.
- Zillmann, D. (2000). Mood management in the context of selective exposure theory. Communication Yearbook, 23, 103-123.
Appendix A. Sub-metric scoring guide #
Each sub-metric is scored on a 0 to 50 scale using the following anchors. A reviewer applies each anchor by asking whether the title clearly clears that threshold; if not, they move down to the next anchor.
| Score | Anchor | Meaning |
|---|---|---|
| 50 | Definitional exemplar | Sets the standard for this sub-metric in the database. Reference titles at 50 include The Wire (Narrative Complexity), Succession (Dialogue Density), Severance (Cognitive Load), Shogun (Production Value), Six Feet Under (Emotional Range). |
| 40 | Exceptional | Demonstrates the sub-metric at a level clearly above the stimulating tier but short of definitional. |
| 30 | Strong | Consistent, intentional, and reliably present. Typical of stimulating-tier titles. |
| 20 | Present | The sub-metric is visible but intermittent or modest in scope. Typical of competent-tier titles. |
| 10 | Minimal | The sub-metric is detectable only occasionally. Typical of passive-tier titles. |
| 0 | Absent | The sub-metric is not present or is actively undermined. |
Reviewers may score at any integer value between anchors. Half-point increments are not used.
Appendix B. Database summary statistics #
As of August 23, 2026.
Total scored titles: 2,606
- Adult titles: 2,379
- Children's titles: 227
Distribution by tier
- Masterclass (160+): 681 (26.1%)
- Stimulating (130-159): 896 (34.4%)
- Competent (100-129): 868 (33.3%)
- Passive (70-99): 142 (5.5%)
- Numbing (<70): 18 (0.7%)
Supplementary dimension coverage
- SEL total displayed: 227 (all children's titles). Current-standard kids scores preserve five competency judgments; legacy SEL totals are reviewed only at the total level pending competency review.
- Cinematic scored: 428 titles where soundtrack is a significant artistic element
- EI (Emotional Intelligence) scored: 0 public. Non-operational research archive (Section 5). Not displayed publicly and no launch is scheduled.
Platform and decade breakdowns are maintained internally and available on request to credentialed researchers.
Appendix C. CASEL framework detail #
For current-standard kids scores, the SEL dimension in Section 4 follows the five CASEL competencies, and each is assessed through the questions below. Legacy SEL values are total-level judgments without a preserved competency receipt; competency-level review remains pending.
Self-Awareness. Does the content show characters identifying their own emotions? Recognizing personal strengths and limitations? Naming feelings accurately? Modeling the development of self-confidence?
Self-Management. Does the content show characters regulating emotions under pressure? Practicing impulse control? Setting goals and working toward them? Demonstrating stress-management techniques appropriate to the developmental stage?
Social Awareness. Does the content show characters taking others' perspectives? Practicing empathy across demographic lines? Recognizing cultural, familial, and individual differences without flattening them? Respecting people unlike themselves?
Relationship Skills. Does the content show characters communicating clearly? Cooperating across difference? Building healthy relationships? Resolving conflicts constructively? Seeking and offering help?
Responsible Decision-Making. Does the content show characters considering consequences? Applying ethical reasoning appropriate to the developmental stage? Evaluating choices? Taking responsibility for actions?
For current-standard scores, each competency is scored on a 0 to 10 subscale. The five subscale scores sum to the 0 to 50 SEL score.
Appendix D. Current editorial accountability #
Final public score authority is human and rests with the named accountable editor. See Jordan Robinson's author page for current biography and credentials.
Education Specialist and Nationally Certified School Psychologist. Shapes and reviews the TVI Kids methodology and personally reviews only pages and TVI Kids Essential designations attributed to her. This role is methodology review, not catalog-wide per-title review.
TVI does not claim a current advisory board, collective expert-panel review, or a third methodology signatory.
Appendix E. Glossary #
CASEL. The Collaborative for Academic, Social, and Emotional Learning, the nonprofit organization that maintains the SEL framework used in K-12 education and in TVI's SEL dimension.
Cinematic Score. A 0 to 50 dimension measuring composition quality, emotional impact, thematic integration, and memorability of a title's score and soundtrack. Reported alongside the IQ Score where applicable. Does not feed into the IQ Score formula.
Cognitive Stimulation (CS). The first dimension of the IQ Score. Assesses the tracking, inference, working-memory, and conceptual demands the work places on the declared reference viewer. Weighted at 40%.
Composite score. A single metric computed from multiple sub-metrics. The IQ Score is a composite of Cognitive Stimulation, Educational Value, and Craft & Quality.
Educational Value (EV). The second dimension of the IQ Score. Assesses accurate content, visible reasoning, emotional-process modeling, usable behavior, and transfer support present in the work. It does not claim that a particular viewer learned or changed. Weighted at 35%.
EI Score (Emotional Intelligence Score). An archived, non-operational research concept described in Section 5. It is not scored or displayed publicly and has no current signatory.
Craft & Quality (CQ). The third dimension of the IQ Score. Measures craft and engagement. Weighted at 25%. Named Entertainment Quality (EQ) through v1.2; renamed in v1.3 to describe the dimension accurately and to retire the abbreviation collision with the Emotional Intelligence dimension, which uses the display label "EI Score", see Section 5.
Inter-rater reliability. The reproducibility of scores produced by independent reviewers applying the same frozen instrument to the same scored object. TVI's reliability program reports absolute agreement, numerical differences, and tier agreement; it is not yet the basis of a public validation claim.
IQ Score. The composite 0 to 200 content-rating score published by TVI. Not a measurement of viewer intelligence.
Masterclass / Stimulating / Competent / Passive / Numbing. The five tier categories defined in Section 7.
SEL. Social-Emotional Learning. A 0 to 50 dimension applied to children's content and scored against the CASEL framework. Does not feed into the IQ Score formula.
Sub-metric. One of the thirteen constituent measures used to produce a dimension score. Cognitive Stimulation and Craft & Quality each contain four sub-metrics scored 0 to 50 and averaged. Educational Value contains five sub-metrics scored 0 to 10 and summed.
Version history #
Entries before the August 19, 2026 correction preserve the historical release record. Names and roles in those rows do not describe current editorial authority. Current accountability is stated in Appendix D.
| Version | Date | Changes | Historical release record |
|---|---|---|---|
| 0.1 | April 19, 2026 | Internal draft. Established dimensional framework, scoring rubric, SEL framework, and integrity protocols. Not distributed. | Robinson, Witty |
| 1.0 | April 23, 2026 | Initial public release. Expanded Section 3 with per-sub-metric peer-reviewed citations. Separated SEL into Section 4 with credentialed-reviewer attribution and anchor scores. Added Section 7 "Supplementary Dimensions (Banked)". Added formal dispute process. Added References and Appendices A-D. | Robinson, Witty |
| 1.1 | April 28, 2026 | Added Advisory Board structure (Appendix D) with three credential blocks. Promoted EQ/EI dimension from "banked" to its own Section 5 (in active development) with full theoretical grounding, five sub-dimension definitions, proposed anchor calibrations, and signatory arrangement with Alexander William Gigler, PsyD as signatory-designate. Added EI/EQ naming clarification. SEL anchor calibrations in Section 4.4 verified against live database. | Robinson, Witty, Gigler (advisory, EI dimension) |
| 1.2 | May 4, 2026 | Founder credential phrasing in Section 1.4 and Appendix D softened from "Master of Public Health concentration in research methodology" to "Master of Public Health, with substantial coursework in research methodology", accurate reflection of the credential. | Robinson, Witty, Gigler (advisory, EI dimension) |
| 1.3 | July 15, 2026 | Third dimension renamed from Entertainment Quality (EQ) to Craft & Quality (CQ) across the document, resolving the divergence between this document and the published methodology page, and retiring the EQ abbreviation collision with the Emotional Intelligence dimension in Section 5. Section 3.2 Educational Value rewritten from four factual sub-metrics to the five-part rubric already in operational use and already referenced by Section 4.5 (Academic Content, Emotional Intelligence, Critical Thinking, Life Skills, Knowledge Transfer), each scored 0 to 10 and summing to the dimension's 0 to 50 score. Section 3 intro, Section 6.2 scoring protocol, and Appendix A updated to the thirteen-sub-metric structure. Two grounding citations added (Immordino-Yang & Damasio 2007; Halpern 1998). No score, weight, threshold, or formula change. | Robinson, Witty, Gigler (advisory, EI dimension) |
| 1.3.1 | August 2, 2026; corrected August 19 and August 23, 2026 | Patch release. Specified exact-half-up rounding for C, Q, and the composite; replaced viewer-outcome language with observable work-property definitions; operationalized Knowledge Transfer as transfer architecture; narrowed credential language to methodology review; and pre-specified the Phase 2 reliability design. The August 19 correction removed stale current-panel and signatory claims, clarified that EI material is a non-operational research archive, and aligned Cordelia Witty's role with the public credential boundary. The August 23 correction aligned the named SEL anchors with the current catalog, disclosed the staged but unexecuted season-range consolidation, scoped five-competency SEL production to current-standard kids scores, corrected current distribution metadata, and completed EI naming across current surfaces. These were correction-only changes; no rubric, weight, threshold, dimension, or composite-formula change. | Robinson |