The context
Verity Goods is a premium DTC beauty brand, strong reviews, strong imagery, persuasive copy. The trust signals were all there, just spread across five surfaces.
In early 2026, support tickets from agent sessions ran 38% higher than from web sessions. Shopping agents couldn't connect delivery, returns, warranty, and ingredients, and pushed buyers to support or to competitors.
Where the path broke
We tested 60 real buyer questions against ChatGPT Shopping, Perplexity, and Claude. Three breaks covered 80% of the failures:
Ingredients as marketing, not data
'Clinical-grade' was a phrase; the actual list sat in an image tab. Agents quoted the claim but not the evidence, which read as unverified.
Delivery promised three ways
The PDP said 3–5 days, the policy page said 5–7, the cart said 2–4 express. Agents showed the most cautious one; a competitor promising two days won the recommendation.
Conditions buried under the headline
'Money-back guarantee' was a banner; the conditions were a 4,000-word PDF. Agents cited half the promise, and the gap became a support ticket.
"Our pages were strong, but only for humans. Now they're strong for agents too, without any loss for customers."
Results after 6 weeks
Three headline numbers, and the main one breaks down.
42%fewer policy escalations
2.4×clearer answer coverage
−42% policy escalations
The 42% broke down into four levers we measured separately:
consistent delivery promises+18%
structured ingredients+14%
citable guarantee conditions+8%
FAQ schema migration+2%
What they kept
The win wasn't a rewritten policy page. It's that one rule has one value, everywhere, so the next product launches consistent, and a CI drift check plus a weekly coverage test keep it that way without anyone maintaining a checklist.