← Back to Wondr Health scales trusted health coaching with Claude

Source proof for Wondr Health scales trusted health coaching with Claude

Source-bound proof

Verified source excerpts for every supported field

Each colour maps a published value to the exact source passage used to support it. Only bounded excerpts are public; administrators can inspect the complete captured source.

15 fields supported

Title

Wondr Health scales trusted health coaching with Claude

derived · high
…ere Ask questions about this page Copy as markdown Case study | Claude Platform Wondr Health scales trusted health coaching with Claude Try Claude Industry: Software Company size: Startup Product: Claude Platform Pa…

Description

Wondr Health, a digital behavior-based weight management program offered as an employer benefit, worked with AI engineering firm Blank Metal to build Wonda, an AI coaching layer running on Claude via Amazon Bedrock: Claude Opus 4.5 powers coaching conversation reasoning, Sonnet runs a parallel safety subagent that monitors every conversation turn and escalates risk to human coaches via Zendesk, and Haiku runs automated evaluations, with onboarding answers grounded in Wondr's curriculum through a RAG pipeline. Before launch, Wondr's clinical and coaching teams tested Wonda against more than 700 questions across six user personas. The AI coaching layer compresses onboarding that historically took a human coach 30 to 60 minutes into a 10-minute guided conversation.

derived · high
…oarding time reduced From 30 to 60 minutes into a 10-minute guided conversation Wondr Health is a digital health company whose behavior-based weight management program is offered to employees as part of their employer-sponsored benefits plan. The program combines a structured curriculum, human coaches, and accountability…
…ting change possible for participants. To bring that experience to more people, Wondr partnered with AI-native engineering firm Blank Metal across a three-sprint engagement to build Wonda, an AI-powered coaching layer that extends Wondr's human coaches rather than replacing them. The coach is powered by Claude through Amazon Bedrock. Its first phase, now built, tested, and heading into a beta rollout, reshapes t…
…etting, and answers grounded in Wondr's curriculum through a RAG pipeline. Claude Opus 4.5 powers the coaching agent's reasoning, Claude Sonnet detects escalation scenarios in the safety subagent, and Claude Haiku runs the automated evaluations as an LLM judge, all through Amazon Bedrock. At every conversation turn, the safety subagent checks the input and conversati…
…e of the most critical moments in the program. With Claude, Wondr Health built: An AI coaching layer that reduces onboarding time from 30 to 60 minutes into a 10-minute guided conversation, without shortening the relationship it leads into A testing console that let clinical and coaching teams put 700+ questions throu…

Company

Wondr Health

classification · high
…ere Ask questions about this page Copy as markdown Case study | Claude Platform Wondr Health scales trusted health coaching with Claude Try Claude Industry: Software Company size: Startup Product: Claude Platform Pa…

Region

North America

classification · high
…oftware Company size: Startup Product: Claude Platform Partner: AWS Blank Metal Location: North America 700+ questions tested by Wondr's team before any real participant met thei…

Industry

Technology & Software

classification · high
…ude Platform Wondr Health scales trusted health coaching with Claude Try Claude Industry: Software Company size: Startup Product: Claude Platform Partner: AWS Blank Metal Locatio…

Problem

Wondr wanted to scale its coaching programs to serve more people while making coaching relationships stronger, starting with onboarding, where the trust-building that historically took a human coach 30 to 60 minutes had to happen in a 10-minute digital window without losing participants who don't find their footing early. Coaches were stretched thin fielding high-volume, repetitive questions while needing to analyze data across systems to understand a single participant's situation.

derived · high
…t, and Wondr's results rest on the coaching relationships that guide them. The company wanted to scale its programs to serve more people while making those relationships stronger, and the place to start was the beginning. Participants arrive at program launch with questions about what they signed up…
…making those relationships stronger, and the place to start was the beginning. Participants arrive at program launch with questions about what they signed up for and how the program will help them, and the trust-building that historically took a human coach 30 to 60 minutes now happens in a 10-minute digital window. Getting that window right matters for everyone: participants who don't find their footing early can drop off before the program has a chance to work, a loss for their health, for the employer paying for the benefit, and for Wondr's retention. Coaches were stretched thin fielding high-volume, repetitive questions while needing to analyze data across systems to understand a single participant's situation. And Wondr's e…

Excerpt shortened for public display.

Solution

Wondr partnered with AI-native engineering firm Blank Metal across a three-sprint engagement to build Wonda, an AI coaching layer powered by Claude through Amazon Bedrock. Claude Opus 4.5 powers the coaching agent's reasoning, Claude Sonnet runs a parallel safety subagent that monitors every conversation turn and escalates risk to human coaches via a Zendesk ticket, and Claude Haiku runs automated evaluations as an LLM judge, with onboarding answers grounded in Wondr's curriculum through a retrieval augmented generation (RAG) pipeline. LangGraph handles orchestration and LangSmith captures full conversation tracing, and a Next.js testing console let clinical and coaching teams put 700+ questions through Wonda across six user personas without engineering support.

derived · high
…etting, and answers grounded in Wondr's curriculum through a RAG pipeline. Claude Opus 4.5 powers the coaching agent's reasoning, Claude Sonnet detects escalation scenarios in the safety subagent, and Claude Haiku runs the automated evaluations as an LLM judge, all through Amazon Bedrock. At every conversation turn, the safety subagent checks the input and conversati…
…iku runs the automated evaluations as an LLM judge, all through Amazon Bedrock. At every conversation turn, the safety subagent checks the input and conversation history for clinical risk, off-topic behavior, and escalation triggers. When it fires, it automatically creates a Zendesk ticket with the risk level and a full conversation summary, gated for human coach follow-up based on severity. An API gateway screens every request, handling authentication, rate limiting, a…
…ate limiting, and participant data filtering before anything reaches the model. LangGraph handles orchestration and LangSmith captures full conversation tracing, so every interaction is logged and inspectable. Custom automated evaluators run verbosity checks and content accuracy scoring.…
…tom automated evaluators run verbosity checks and content accuracy scoring. And a Next.js testing console on Vercel gave Wondr's clinical and coaching teams direct access to test Wonda's responses across six user personas, each built with realistic simulated metadata, without needing engineering support. Wondr accelerated the build by showing up prepared. The company came in with us…

Business value

Wonda delivers onboarding in a 10-minute conversation, compressing what historically took a coach 30 to 60 minutes and letting participants get to human coaching more efficiently. Between the two human evaluation rounds, Wonda's scored response quality kept improving, with strong reception on tone, empathy, and personalization from the coaching and clinical teams, and every repetitive onboarding question Wonda takes on is meant to free a coach for conversations that require a human.

derived · high
…lead, Blank Metal The outcome Measured quality and trust earned on both fronts Wonda delivers onboarding in a 10-minute conversation, compressing what historically took a coach 30 to 60 minutes, now allowing participants to get to human coaching more efficiently. Between the two human evaluation rounds, Wonda's scored response quality k…
…0 minutes, now allowing participants to get to human coaching more efficiently. Between the two human evaluation rounds, Wonda's scored response quality kept improving, with strong reception on tone, empathy, and personalization from the coaching and clinical teams, the two groups who needed to trust Wonda most. The deeper result echoes what the Zendesk data first revealed: evaluators told…
…nput into Wondr's care model. The design also keeps coaches at the center: every repetitive onboarding question Wonda takes on is meant to free a coach for the conversations that require a human. The coaches' expertise, empathy, and clinical judgment remain the core of…

AI capabilities

Conversational AI, Agentic AI

classification · high
…cing the modernization work Wondr's environment would need for production. Claude sits at the center of an agentic architecture handling participant conversations: personalized onboarding, goal setting, and answers grounded in Wondr's curriculum through a RAG pipeline. Claude Opus 4.5 powers the coaching agent's reasoning, Claude Sonnet detec…

Business functions

Customer Service & Support

classification · medium
…ealth, for the employer paying for the benefit, and for Wondr's retention. Coaches were stretched thin fielding high-volume, repetitive questions while needing to analyze data across systems to understand a single participant's situation. And Wondr's existing Zendesk bot had surfaced something unexpected: partic…

Technology

Claude Platform, Amazon Bedrock, LangGraph, LangSmith, Zendesk

classification · high
…etting, and answers grounded in Wondr's curriculum through a RAG pipeline. Claude Opus 4.5 powers the coaching agent's reasoning, Claude Sonnet detects escalation scenarios in the safety subagent, and Claude Haiku runs the automated evaluations as an LLM judge, all through Amazon Bedrock. At every conversation turn, the safety subagent checks the input and conversati…
…ate limiting, and participant data filtering before anything reaches the model. LangGraph handles orchestration and LangSmith captures full conversation tracing, so every interaction is logged and inspectable. Custom automated evaluators ru…

Deployment model

Cloud (Amazon Bedrock)

classification · high
…ching layer that extends Wondr's human coaches rather than replacing them. The coach is powered by Claude through Amazon Bedrock. Its first phase, now built, tested, and heading into a beta rollout, reshapes t…

Deployment options

cloud

classification · high
…ching layer that extends Wondr's human coaches rather than replacing them. The coach is powered by Claude through Amazon Bedrock. Its first phase, now built, tested, and heading into a beta rollout, reshapes t…

Headline outcome

derived · high
…e of the most critical moments in the program. With Claude, Wondr Health built: An AI coaching layer that reduces onboarding time from 30 to 60 minutes into a 10-minute guided conversation, without shortening the relationship it leads into A testing console that let clinical and coaching teams put 700+ questions throu…

Use case type

Conversational assistant

classification · high
…cing the modernization work Wondr's environment would need for production. Claude sits at the center of an agentic architecture handling participant conversations: personalized onboarding, goal setting, and answers grounded in Wondr's curriculum through a RAG pipeline. Claude Opus 4.5 powers the coaching agent's reasoning, Claude Sonnet detec…
Capture details
Captured
25 Sept 2026, 07:11 UTC
Extractor
fetch-strip@1
Snapshot hash
e6b8c4935211be7a5714809521389910edfa64f04d087084566a8c99e99c2e25