Partner Content

Artificial Intelligence and Machine Learning

July 10, 2026

4 min read

Why the Model Behind Your Synthetic Research Tool Matters

Why the Model Behind Your Synthetic Research Tool Matters

Learn how to evaluate synthetic research tools and build confidence in AI-generated data for better business decisions.

The market raced to launch synthetic research tools. Qualtrics took a different path,and the difference shows up in the data.

The benefits of synthetic research are  real: faster turnaround, lower cost, no recruiting timelines, but as more insights teams incorporate AI-generated data into their workflows, critical questions have to be answered. Which kind of model should your team be using? What is the model actually doing when it generates a response and how are those responses calibrated? Here’s why the answers matter.

Where General-Purpose AI Models Fall Short in Synthetic Research

A general-purpose LLM—the technology behind tools most people know as ChatGPT, Gemini, or Claude—generates responses by predicting what text is most probable given a particular input. When you ask one to simulate a respondent ("respond like a 28-year-old male who works in manufacturing in the midwest"), it produces answers that reflect the cultural patterns and linguistic norms baked into its training data, which for general-purpose models is largely derived from the internet. The output reads as plausible, and for qualitative exploration and early hypothesis generation, it can be genuinely useful. The problem surfaces when researchers attempt to use that output for quantitative analysis.

Real survey populations produce data with natural variance. Inconsistencies between stated and derived preferences, non-linear response patterns, the kind of distributional noise that actually reflects how humans make decisions under uncertainty, are present. General-purpose models tend to flatten that variance, producing responses that cluster around what the persona "typically" believes rather than distributing the way a real sample would. The result looks structured at the surface, but it often isn’t.

When those outputs are run through more advanced research methods like factor analysis, clustering, key drivers analysis, and conjoint, the data frequently lacks the structural integrity those methods require. These general-purpose models were built to draft emails, summarize documents, and answer open-ended questions. They were not designed to simulate the statistical behavior of human survey respondents. Asking them to is a category mismatch, and the flaws tend to become visible exactly when the analysis becomes the  most consequential.

Three Common Approaches to Synthetic Data Generation and Their Tradeoffs

Most synthetic research tools in the market today fall into one of three, nonequivalent categories:

Prompt-engineered personas instruct a general-purpose LLM to respond like a particular audience segment through system and user-level instructions. These tools have genuine utility for early qualitative exploration and hypothesis generation. Where they fall short is in quantitative simulation. Without deeper model customization, responses cluster around cultural archetypes rather than reflecting the authentic variance of a real population. Data that looks structured at first doesn't survive statistical scrutiny.

RAG (Retrieval-Augmented Generation)-based digital twins go a step further by grounding the model in a curated database of documents, interview transcripts, or prior research data. This produces more varied and realistic-looking outputs and is particularly useful for niche audiences where real human data can anchor responses more precisely. The limitation is that any bias present in the retrieval database carries through to the synthetic output. Like prompt engineering, RAG adds a layer of specificity without modifying the base model itself.

Fine-tuned LLMs are built differently. Rather than shaping a general-purpose model through instructions or supplementary data, fine-tuning retrains it on real human survey behavior—changing how it generates responses fundamentally. The result isn't a model that simulates what a persona would say, it's a model that learns to reproduce the statistical behavior of real respondents.

Qualtrics Synthetic Model Approach

Fine-tuning is the investment Qualtrics made before bringing synthetic audiences to customers. Rather than adapting a general-purpose model through prompting or RAG alone, the team developed a foundational model trained specifically on anonymized market research data, optimized to generate survey responses that hold up under rigorous statistical scrutiny. Independent testing showed the Qualtrics model is 12x more accurate than general-purpose LLMs in predicting human survey responses.

Fine-tuning forms the core of the architecture—working alongside the other methods, not in place of them. From there, prompt engineering guides how that model interacts with each survey instrument, maintaining appropriate context and persona characteristics throughout data collection. RAG can be deployed to ground outputs in current market information rather than relying solely on historical training data. Each layer addresses a different dimension of the problem.

For insights leaders managing pressure to move faster without sacrificing the rigor their credibility depends on, the quality of the underlying model matters more than the interface built on top of it.

Synthetic data that can't survive statistical scrutiny won’t accelerate research timelines in any meaningful sense. The reckoning just moves downstream, adding risk to already complex decisions. Investment in foundational model development is the harder path, but also what makes the output trustworthy enough to act on.

Large Language Models (LLMs)artificial intelligencesynthetic data

Comments

Comments are moderated to ensure respect towards the author and to prevent spam or self-promotion. Your comment may be edited, rejected, or approved based on these criteria. By commenting, you accept these terms and take responsibility for your contributions.

Derrick McLean, PhD

Derrick McLean, PhD

Product Scientist, Edge COE at Qualtrics

3 articles

author bio

Disclaimer

The views, opinions, data, and methodologies expressed above are those of the contributor(s) and do not necessarily reflect or represent the official policies, positions, or beliefs of Greenbook.

About partner

Qualtrics builds technology that closes experience gaps, transforming insight into impact.

More from Derrick McLean, PhD

Five Questions To Ask When Choosing Your Synthetic Research Partner
Artificial Intelligence and Machine Learning

Partner Content

Five Questions To Ask When Choosing Your Synthetic Research Partner

Learn how to evaluate synthetic research tools with 5 key questions on data generation, model validation, and human insight.

Testing Synthetic Data Against Academic Benchmarks: A Replication Study
Data Science

Partner Content

Testing Synthetic Data Against Academic Benchmarks: A Replication Study

7 min read

Qualtrics examines how synthetic data performs against academic benchmarks, addressing trust and validation gaps in AI-driven research.

How to Be Memorable When AI Is Churning Out Blah Blah Content
Artificial Intelligence and Machine Learning

How to Be Memorable When AI Is Churning Out Blah Blah Content

Discover why a unique point of view is essential in the insights industry. Learn how to make your research stand out amidst AI-generated content.

Iosetta Santini

Iosetta Santini

Account Director at Keen as Mustard Marketing

Lost in Translation: The Reality of Synthetic Research Tools in Japan and Korea
Artificial Intelligence and Machine Learning

Partner Content

Lost in Translation: The Reality of Synthetic Research Tools in Japan and Korea

Survey data reveals how insights professionals in Japan and South Korea perceive, use, and plan to adopt synthetic research.

Ryan Kaneko

Ryan Kaneko

Vice President, Global API Alliance at GMO Research & AI, Inc.

What I’ve Learned Co-Hosting the MRII Podcast: AI Presents A Unique Opportunity for Insights to Reinvent Itself
Executive Insights

What I’ve Learned Co-Hosting the MRII Podcast: AI Presents A Unique Opportunity for Insights to Reinvent Itself

Explore why AI is an opportunity for insights teams to reinvent their role, increase business impact, and improve decision-making.

Nick Graham

Nick Graham

Founder at Vertemis

Synthetic Respondents Explained: What They Are, How They Work, and When to Trust Them
Artificial Intelligence and Machine Learning

Synthetic Respondents Explained: What They Are, How They Work, and When to Trust Them

7 min read

Synthetic respondents use AI to simulate survey participants. Learn how they work, when they're accurate, and when real respondents are still essentia...

Ashley Shedlock

Ashley Shedlock

Content Producer at Greenbook

Sign Up for
Updates

Get content that matters, written by top insights industry experts, delivered right to your inbox.

67k+ subscribers