Skip to main content
Case Study · Enterprise Energy

More Experiments.
The Same Evidence Bar.

An in-house experimentation program across 5 regulated energy retail brands—Reliant Energy, Direct Energy, Green Mountain Energy, Cirro Energy, and Discount Power—with shared standards for prioritization, QA, analysis, and financial modeling.

CompanyNRG Energy
RoleCRO & UX Manager
Period2023 – Present
IndustryRegulated Energy
$30M+
Impact in NRG Internal Program Reporting
20→100+
Annual Experiment Velocity
100+
Experiments Run in 2025
~40%
Internal Workflow-Time Estimate
Internal 2025 program reporting recorded $30M+ in impact across the portfolio
Annual throughput scaled from roughly 20 experiments to 100+ by 2025
A selected Reliant readout recorded a 12% lift in enrollment confirmations
A bundled Direct Energy variant recorded a 60% mobile enrollment lift
A Green Mountain Energy readout recorded roughly 3× call sales with incomplete call tracking
Test-period outcomes were translated into modeled annual impact with assumptions disclosed
Briefs, QA, power checks, guardrails, and decision rules remained mandatory as volume grew
The Challenge

Five Brands. Zero
Shared Methodology.

NRG Energy operates multiple regulated energy retail brands, each with different customer bases, regulatory environments, and acquisition funnels. When I joined, experimentation was fragmented—roughly 20 tests per year, with inconsistent standards for connecting test evidence to revenue models.

The central business question was: what evidence supports the value of the experimentation program? Making test-period outcomes, assumptions, and uncertainty visible gave finance and leadership a more reviewable basis for capacity and investment decisions.

The Journey

From Foundation
to Compound Growth

Foundation · Q1–Q2 2023

Consolidating a Fragmented Program

Inherited testing across 5 brands without one shared operating method. Reviewed the existing portfolio, documented measurement gaps, and built a hypothesis-to-revenue evidence chain. The operating standard included minimum detectable effect, power calculations, documented stopping rules, and guardrail checks.

Scale · 2023–2025

From Roughly 20 to 100+ Tests by 2025

Rolled out shared intake, prioritization, QA, and readout standards across Reliant Energy, Direct Energy, Green Mountain Energy, Cirro Energy, and Discount Power. Modeled impact made assumptions reviewable alongside the experimental evidence; it did not turn every metric movement into realized revenue.

Compound · 2024 – 2025

AI-Augmented Experimentation

Added approved AI-assisted steps for candidate-pattern review, analysis support, and drafting. A first-person, uncontrolled workflow estimate moved from roughly eight to five hours per readout; statistical checks, source-data validation, and human review remained mandatory. The estimate does not isolate AI as the cause.

Inside the Program

Inside the NRG Experiment Portfolio

Test-level details stay internal—but the themes show how the program operated. The documented standard required a named behavioral mechanism, pre-specified primary outcome, power and minimum-detectable-effect checks, guardrails, and an approved stopping and readout procedure. Those controls make the evidence reviewable; they do not guarantee a positive result.

Friction & cognitive load

Reliant Enrollment-Flow Simplification

A selected internal readout recorded a 12% lift in enrollment confirmations, about 100 additional confirmations over 46 days, and approximately $299K in modeled annual impact. The bundled variant did not isolate one causal element.

Visual hierarchy & comprehension

Direct Energy Home-Bundle Experience

A bundled copy and visual-hierarchy variant recorded a 60% mobile enrollment lift, 70 additional enrollments over 36 days, and approximately $177K in modeled annual impact. The readout did not isolate copy from layout.

Channel preference & salience

Green Mountain Energy Call-Channel Capture

A selected internal readout recorded roughly 3× call sales and 948 additional call sales over 57 days after a bundled change, with approximately $523K in modeled annual impact. Incomplete call tracking means the movement is not perfectly isolated incremental revenue.

Visibility & information hierarchy

Green Mountain Energy Hero Layout

A selected internal readout recorded a 7% secondary lift in enrollment starts, a 35% change in the chart-to-enrollment ratio, and approximately $212K in modeled annual impact. The secondary ratio was diagnostic, not proof of higher-quality users.

Friction & truthful benefit framing

Credit and Deposit-Step Friction

Internal tests evaluated identity-verification framing and the salience of lower-commitment options. Their readouts were interpreted as funnel-specific evidence, not a reusable conversion or cost-savings benchmark.

Relevance & behavioral targeting

Personalization via CDP Segmentation

A bundled Tealium and Optimizely treatment recorded a 23% lift in lead acquisition versus control in an internal readout. The result does not isolate AI, segmentation, content, or delivery tooling as the sole cause.

Every test moves through the same 12-step operating pipeline — from idea intake to executive readout and handoff.See the pipeline →

Behavioral Science in Action

The Principles
Behind the Numbers

The NRG program used named behavioral mechanisms where customer evidence supported them. These four examples generated testable hypotheses; none was treated as a guaranteed outcome.

Loss Aversion

Used truthful loss-versus-gain framing as a hypothesis for rate-plan messaging, then evaluated the result within each brand instead of assuming the mechanism would transfer.

Choice Architecture

Tested whether fewer, better-explained plan choices and clearly disclosed defaults improved qualified enrollment decisions.

Social Proof

Tested substantiated local choice cues on rate-comparison pages. One internal readout recorded a 22% shorter time-to-decision during its measurement window.

Anchoring

Tested plan order as an anchoring hypothesis while retaining average revenue per customer as a downstream guardrail.

Methodology

Built on the
PRISM Method

The NRG experimentation program uses PRISM to frame a pre-test value range, document its assumptions, and compare the post-test evidence with that range. The result retains its uncertainty; a metric movement is not automatically booked as realized revenue.

See the PRISM Method →
Probe — Behavioral diagnosis
Revenue Rank — Impact prioritization
Implement — Hypothesis & execution
Score — Revenue measurement
Multiply — Compound & scale
Continue Reading

See all case studies

Explore the geo-experimentation work at Silicon Valley Bank, see individual test-level evidence in the Experiments hub, or read weekly breakdowns on Lean Experiments.

SVB Case Study → Enterprise CRO program → Experiment Case Studies → All Work
Better Decisions Newsletter

Make Growth Decisions
You Can Defend

Know what to test, when to trust the result, and what to do next. Practical decision guides for analysts, growth teams, and founders.