Brand Voice in the AI Era

Tone vs Voice: The Difference That Trips Teams Up

Voice stays fixed. Tone moves with context. Most \"off-brand\" feedback confuses the two, and the fix that follows usually breaks the wrong thing.

Most "this doesn't sound like us" feedback isn't about voice. It's about tone, misdiagnosed. Someone reads a launch post that feels flat, or an outage notice that feels chirpy, and calls the whole thing off-brand. The fix that follows usually touches the wrong layer, and a team ends up rewriting a style guide over a single badly-toned email.

This piece works from years of building and auditing brand voice systems for content teams who kept having the same argument every quarter without naming what the argument was actually about.

The split

Voice is what stays the same. Tone is what's supposed to move.

Voice covers vocabulary, sentence rhythm, and point of view, the stuff a brand would never say out loud. Strip the logo off a page and a reader should still place it. Tone sits on top and adjusts for the moment, warmer for onboarding, sober for an outage, sharper for a launch. The words don't change. The temperature does.

Run one message through two tones without touching the voice underneath. For an outage: "We're on it. Fix coming, no ETA yet, we'll update you the second we know more." For onboarding: "We're on it, and we'll walk you through what's happening as soon as we know more, so you're not left guessing." Both stay short, neither overpromises, and the same person is talking throughout. What moved was warmth, not identity.

Here's the split laid flat:

  • Voice is fixed across every channel and every mood. It covers vocabulary, sentence rhythm, point of view.
  • Tone is adjusted per message, per reader state. It covers warmth, formality, pace, energy.
  • Voice answers "who is talking." Tone answers "how do they feel about this specific thing right now."

Microsoft's own brand voice guidance draws this line for a much bigger organisation than most of us run. Its style guide states plainly that voice stays constant while tone shifts, moving "from serious to empathetic to lighthearted" depending on the reader's state of mind. Salesforce's brand voice guide, in its own published materials, frames it almost identically. Voice is the constant "true north." Tone adapts to audience, medium, and situation, and the practical direction underneath that is to keep customer-facing copy direct, concise, and helpful. Grammarly's brand voice guidance goes further and treats tone as a subset of voice, shaped by word choice and shifting circumstance rather than a separate system entirely. Different framing, same underlying split.

Why teams blur it

Most style guides file this under one heading, "voice and tone," and that's the moment the trouble starts. Once both concepts share a section, they start sharing rules.

A reviewer flags a launch announcement as too stiff. The fix applied is a permanent change to the brand's vocabulary, because nobody separated "this needed more energy for this moment" from "this doesn't sound like us." Weeks later an outage notice comes back too breezy, and instead of tightening the tone for that one message, someone rewrites the whole voice toward formality. Neither fix touches the actual problem. Each one drags the baseline further from where it started.

Say a small team is producing a product update, a support macro, and a social caption in the same week, written by three different people. Without a documented split between what's fixed and what's expected to flex, each writer guesses at both, and the guessing compounds. Nobody can point to when the voice actually changed, only that it has.

Mailchimp went the other way on purpose. Its content style guide separates voice from tone explicitly, describing voice as the stable brand personality and tone as the layer that flexes by situation, and it leans on real examples so teams apply the split the same way across channels rather than re-litigating it per piece. That specificity is what most internal guides skip. They say "adjust tone as needed" without ever showing what "as needed" looks like on the page.

Voice stays put

Voice shouldn't change for channel, campaign, or mood. If the brand is direct and plain-spoken, it stays direct and plain-spoken in the apology email, the pricing page, and the meme on social. What changes around it is warmth, formality, and pace, never the underlying vocabulary or point of view.

Teams get this backwards most often during a big launch. Voice bends toward hype because the moment feels like it calls for it, and regular readers say the brand felt off that week. Salesforce's own 2026 State of Agentic Marketing survey found that brand voice consistency was flagged as a real concern by 39.1% of respondents, which tracks with what happens once AI-assisted drafting scales output across more writers and more channels without anyone holding the line. The guardrails that keep a voice steady matter most in exactly these high-pressure moments, when the temptation to abandon the usual register is strongest.

Tone moves with context

Tone should shift every time the reader's situation changes, and it should shift by a lot. Onboarding needs warmth and patience, because the reader is unsure and wants reassurance more than efficiency. A launch needs confidence and some real energy, because the reader is already primed to be excited. An outage needs the opposite, measured and transparent, no forced cheerfulness, because chipper copy on a broken product reads as tone-deaf rather than on-brand. Pricing needs a flat, low-pressure register, since any hint of salesmanship reads as a red flag to someone comparing costs. Legal notices need precision and nothing else, because the reader is checking for exact meaning, not personality. Support needs to match the specific complaint in front of it, calm for confusion, direct for anger, brief for anything routine.

Put two examples from the same voice side by side. An outage notice: "Some of you can't log in right now. We know, we're already in it, and we'll post an update within the hour whether we've fixed it or not." A launch post: "This has been months in the making and it's finally live. Go poke at it, we built it for exactly the complaints you've been sending us." The directness hasn't moved. One is calm because the reader is anxious. The other is energetic because the reader is already on board.

Intercom's own guidance for its AI support agent, Fin, works the same way operationally: teams define how it speaks to match brand voice, then layer tone-of-voice settings on top for the specific conversation, treating the two as separate configuration steps rather than one blended setting.

Gucci's use of AI-assisted support is a sharper version of the same idea. Salesforce's Customer 360 case study describes how Gucci trained models on past Gucci communications so replies to client advisors came out sounding "Guccified," meaning the voice held constant while the actual wording adapted to whatever the client was asking about. The voice didn't get luxurious for luxury questions. It was already luxurious. The tone adjusted to the situation.

Brief them separately

Voice gets briefed once, in a standing document every writer works from. It shouldn't need re-explaining per assignment any more than the product's name does. Defining a brand voice properly, up front, is what makes that possible. HubSpot's brand style guidance points to this exact contrast when it cites Slack's voice descriptors and Mailchimp's framework as examples of companies that operationalised the split rather than leaving it as a vague aspiration.

Here's what the two briefs look like side by side for the same message, an outage notice going out to customers.

Voice brief (standing, doesn't change): "Direct, plain-spoken, no corporate hedging. Short sentences. First person plural. Never say 'we apologise for any inconvenience,' say what happened and what we're doing."

Tone brief (written per message): "Measured, transparent, no forced cheerfulness. This is an active outage. Acknowledge it plainly, give a real timeframe or say there isn't one, no jokes."

A writer working from both produces the outage line from earlier without inventing anything. Someone handed only the voice brief might write something too breezy for the moment. Someone handed only the tone brief, with no standing voice document, might get the temperature right and the vocabulary wrong, hedging where the brand normally doesn't. Both briefs are needed, and they answer different questions.

A stable voice definition also helps with something less obvious: how AI answer engines summarise a brand. A system pulling from ten different articles surfaces a more accurate picture when the same three or four words describe the voice every time, rather than a slightly different gloss per post. That consistency is part of the ground generative engine optimisation sits on.

Quick reference for applying the split in review:

  • Voice rules cover vocabulary, sentence length, point of view, and what the brand would never say.
  • Tone rules cover warmth, formality, urgency, and energy, and they get written fresh per message.
  • Channel examples run from measured for an outage, to confident for a launch, patient for onboarding, flat for pricing, precise for legal, and matched to the complaint for support.

Support teams run into this daily without naming it. A macro answering "how do I export my data" can stay warm and unhurried. A macro answering "you've been charged twice" needs to drop straight into acknowledgement and next steps, cheerfulness dropped entirely. Same voice, two tones, no rewrite of the macro library required.

A solo founder writing everything alone carries voice and tone in one head already. The documentation buys less there than it does for a five-person team handing assignments back and forth across a shared publishing calendar.

Less work, more on-brand content

Austen runs this whole workflow for you: from research to on-brand drafts that get found by Google and AI.

Start free

More in Brand Voice in the AI Era