The Wren Journal
ProductNews

16× Cheaper, 37% Faster: Wren AI Cloud Now Defaults to GPT-5.6 Luna

Wren AI Cloud is now 16× cheaper and 37% faster per task, with accuracy holding at 95% of evaluation assertions passed — GPT-5.6 Luna is the new default model. Here's what we measured before flipping the switch.

Wren AI Product Team

Wren AI Product Team

Updated: Sep 01, 2026
Published: Sep 01, 2026

16× Cheaper, 37% Faster: Wren AI Cloud Now Defaults to GPT-5.6 Luna

Same trusted answers — now 16× cheaper and 37% faster per task. Here's what we measured before flipping the switch.

What changed

We moved Wren AI Cloud's default model from GPT-5.4 to GPT-5.6 Luna. Nothing changes in how you use Wren — same questions, same connectors, same governed SQL behind every answer. The difference shows up on your bill and your clock. Every Cloud account is on Luna now, with no action needed on your side. (If you self-host Wren AI, this doesn't affect you — you stay in control of which model your deployment runs.)

Before switching, we replayed our full agent evaluation: 22 real ecommerce tasks, from single-fact lookups to multi-turn drill-downs and full GenBI dashboards. Luna handled all of them at a fraction of the cost, and reproduced our previous model's answers to the cent on the core data.

Cost per task · averaged across all 22 evaluation cases

PreviousGPT-5.4
LunaGPT-5.6
$0.036
16× cheaper per task
Every one of the 22 cases came out cheaper — no exceptions.

The numbers

Two things compound. Luna's tokens list at roughly one-twelfth the price to begin with, and Luna reaches the same answer while generating about half as many of the expensive output tokens. The list price does most of the work; trimming output tokens does the rest — together they turn a lower sticker price into a far wider real-world gap.

Cost

Spend per task

$0.036
−94% vs previous

Token price

List price, per million output tokens

$1.20
−92% input drops the same 12.5×

Speed

Model time per task

102.5s
−37% vs previous

Efficiency

Tokens generated per task

6,902
−51% vs previous

Across the whole run, the 22 tasks that cost $12.86 on the old default came in at $0.80 on Luna — a saving of $12.06 on the same workload.

Accuracy held

Cheaper and faster only matters if the answers stay right, with the SQL to prove them. On our evaluation, Luna passed 204 of 215 assertions (95%) and reproduced the previous model's figures to the cent on the core data tasks — total revenue, GMV series, Pareto splits, forecasts and multi-turn reconciliations all landed on the reference values.

95%
of evaluation assertions passed on Luna
22 / 22
tasks came out cheaper — every case, no exception
to the cent
core data answers matched the previous model exactly

You don't need to do anything. Luna is live as the default across all Wren AI Cloud accounts today. Ask your usual questions — you'll just spend less and wait less to get governed answers back.

Wren AI Cloud on GPT-5.6 Luna: 16× lower cost per task, 37% less time per task, 51% fewer tokens per task, with accuracy held at 95% of assertions passed
Wren AI Cloud on GPT-5.6 Luna: 16× lower cost per task, 37% less time per task, 51% fewer tokens per task, with accuracy held at 95% of assertions passed

Same questions. Answered for less.

Give every team trusted answers from the data they already own — now on GPT-5.6 Luna by default. Try Wren AI Cloud free or request a demo.

The Wren Journal

Get the next deep dive in your inbox

Practical GenBI guides, product updates, and customer lessons from the Wren AI team. A couple of emails a month — no noise.

Keep reading

You Can't Trust an AI Agent You Can't Debug.
Insight · Product

You Can't Trust an AI Agent You Can't Debug.

An AI agent that answers business questions has to be debuggable and measurable, or 'earned trust' is just a slogan. How thread tracing, benchmarks, and the AI Advisor close the loop.

July 14, 2026Read