# ConvApparel: Measuring and bridging the realism gap in user simulators

DevFeed: [ConvApparel: Measuring and bridging the realism gap in user simulators](<https://devfeed.tech/articles/convapparel-measuring-and-bridging-the-realism-gap-in-user-simulators-6756.md>)

Original publisher: [Read original article](<https://research.google/blog/convapparel-measuring-and-bridging-the-realism-gap-in-user-simulators/>)

Published: 2026-04-09T11:22:00Z

Content type: article

Language: en

Sources: [The latest research from Google](<https://devfeed.tech/sources/the-latest-research-from-google.md>)

Topics: [Conversational AI](<https://devfeed.tech/topics/conversational-ai.md>), [Generative AI](<https://devfeed.tech/topics/generative-ai.md>), [LLMs](<https://devfeed.tech/topics/llms.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [dataset](<https://devfeed.tech/topics/dataset.md>), [AI Research](<https://devfeed.tech/topics/ai-research.md>), [Statistics](<https://devfeed.tech/topics/statistics.md>), [data](<https://devfeed.tech/topics/data.md>), [Google](<https://devfeed.tech/topics/google.md>)

Tags: [agents](<https://devfeed.tech/tags/agents.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-agents](<https://devfeed.tech/tags/ai-agents.md>), [ai-research](<https://devfeed.tech/tags/ai-research.md>), [conversational-ai](<https://devfeed.tech/tags/conversational-ai.md>), [data](<https://devfeed.tech/tags/data.md>), [evaluation](<https://devfeed.tech/tags/evaluation.md>), [generative](<https://devfeed.tech/tags/generative.md>), [generative-ai](<https://devfeed.tech/tags/generative-ai.md>), [llm](<https://devfeed.tech/tags/llm.md>), [machine-intelligence](<https://devfeed.tech/tags/machine-intelligence.md>), [natural-language-processing](<https://devfeed.tech/tags/natural-language-processing.md>), [research](<https://devfeed.tech/tags/research.md>), [testing](<https://devfeed.tech/tags/testing.md>), [training](<https://devfeed.tech/tags/training.md>), [validation](<https://devfeed.tech/tags/validation.md>)

## AI overview

Google Research introduces ConvApparel, a human-AI conversation dataset and evaluation framework for measuring the realism gap in LLM-based user simulators. It uses Good and Bad agents and validates results through population-level statistics, human-likeness scoring, and counterfactual validation.

## Source excerpt

Generative AI