ai-values-brand-comms-2

Do AI Systems Really Understand What We Value?

A New Study Reveals Why Reward Models Deserve More Scrutiny.
AI tools like ChatGPT are quickly becoming essential for businesses, shaping how we communicate and engage with our communities. At the core of these AI systems lie reward models, unseen judges that evaluate AI-generated responses based on human preferences and values. But how accurately do these invisible decision-makers grasp our intentions and ethics?

A recent study from the University of Oxford, Reward Model Interpretability via Optimal and Pessimal Tokens, invites us into a deeper understanding of this important question. Researchers found that reward models, often perceived as neutral arbiters, can misinterpret human values significantly. For purpose-driven brands, this revelation highlights the need for a more reflective and intentional relationship with technology.

ai values brand comms

The Illusion of Objectivity

The Oxford researchers examined ten popular reward models and uncovered considerable differences even among those designed with similar goals. When faced with identical prompts, each model provided vastly different responses, highlighting the nuanced complexity of encoding human values into algorithms.

For brands committed to building trust and genuine connection, understanding the variability and limitations of these tools is crucial. Blind reliance on AI can inadvertently lead to inconsistencies that compromise your message and impact.

The Power of Language Framing

The study revealed the substantial influence of subtle shifts in wording. Even minor changes in language dramatically influenced AI responses, echoing human cognitive biases. This insight reminds us that language holds powerful sway, whether communication is human or automated.

Organisations dedicated to thoughtful communication must pay close attention to how they frame their interactions with AI, ensuring their true intent and empathy shine through clearly.

Hidden Bias in Identity-Related Language

One of the most critical findings was the presence of subtle yet consistent biases towards particular identity groups. Efforts to ensure safety and harmlessness sometimes unintentionally caused reward models to undervalue or avoid terms associated with marginalised communities.

Organisations striving for inclusion and equity must proactively identify and address these hidden biases. True empathy and understanding require vigilance to ensure every voice is valued and authentically represented.

Beyond Blandness to Authentic Expression

The research also showed that reward models strongly favour common and predictable words. The result is communication that often feels generic, lacking creativity, depth, and distinctive resonance.

For organisations seeking to inspire and lead change, embracing creativity, wonder, and nuanced expression is essential. Avoiding generic communication helps your message resonate authentically, capturing hearts and minds.

Misalignment Propagates Through the System

Because these reward models directly shape the training of larger AI systems, any bias or misalignment at this level can quietly influence the final behaviour of the tools we use. This means that flawed value judgments can ripple through everything from customer interactions to decision support systems.

Human Values Are Complex

Ultimately, the study underscores the complexity of human values. Attempting to simplify them into numeric scores risks oversimplification and distortion.

Purpose-led organisations must honour this complexity. While AI offers powerful capabilities, it should complement—not replace—the uniquely human capacity for reflective, ethical decision-making and visionary communication.

What This Means for Purpose-Led Organisations

To ensure your use of AI aligns authentically with your values:
What to Do Why It Matters
Be Intentional with Language Framing Reward models—and people—respond differently to small shifts in phrasing. Frame messages with care and intention, not just for attention.
Audit AI-Assisted Content for Bias and Blandness AI content tends to be safe and generic. Review it closely to ensure your brand voice remains authentic, inclusive and relevant.
Protect Your Brand’s Human Voice Don’t let automation erode your personality. Keep tone, storytelling and lived experience at the heart of your communication.
Understand the Tools Behind Your Platforms Social and ad platforms use algorithms that influence visibility. Knowing how they work helps you stay intentional and aligned.
Educate and Empower Your Team AI ethics isn’t a solo job. Equip your team with the awareness and confidence to use tools responsibly and reflectively.
Our guiding belief remains steadfast: business should amplify humanity, not diminish it. As we increasingly integrate AI into our communications and operations, let’s remain mindful and curious, always seeking clarity and alignment with the deeper truths that guide us.
Author
Luke Burrell
Mezzanine - Director | Creative Director
Luke Burrell is the Founder and Director of Mezzanine, a Newcastle-based brand strategy and creative consultancy. For more than 25 years, Luke has helped founders, leaders and organisations connect their purpose, values and culture with clear brand strategy and communication. Bringing together creativity, technology and practical action, he helps organisations solve complex challenges, strengthen alignment and build brands designed to endure. Having partnered with start-ups, universities, government bodies and national and international organisations, Luke writes about conscious branding, leadership, organisational culture, purpose and innovation, sharing practical ideas that help organisations create meaningful impact and long-term value.

Sign up to our 'Brand Insights' email

We know how hard it can be to stay aligned, consistent and strategic when managing your brand day-to-day. We provide updates, tips and advice straight to your inbox to help you ensure your brand builds long-term value for your organisation and has a positive impact.