Rosenverse
Human vs. machine: Testing AI’s ability to synthesize and analyze research

This video is only accessible to Gold members. Log in or register for a free Gold Trial Account to watch.

Log in Register

Most conference talks are accessible to Gold members, while community videos are generally available to all logged-in members.

Human vs. machine: Testing AI’s ability to synthesize and analyze research

Gold
Wednesday, March 11, 2026 • Advancing Research 2026

This video is featured in the AI Trends in User Research playlist.

Share the love for this talk
Human vs. machine: Testing AI’s ability to synthesize and analyze research
Speakers: Laura Klein
Link:

Summary

Nielsen Norman Group (NNG) has conducted and continues to conduct extensive research testing various large language model (LLM) tools designed for research synthesis and analysis. Our goal was to determine whether these AI-powered tools could meaningfully accelerate the work of experienced UX researchers. Through rigorous testing across multiple models and specialized research tools, we’ve found that while a few tools provide modest speed improvements for experienced researchers, none come close to replacing human expertise in research synthesis and analysis. The core problem is that these tools consistently exhibit critical flaws: they hallucinate findings, fail to identify meaningful patterns in qualitative data, cannot adequately consider nuanced research questions, and produce only superficial, high-level summaries of participant behavior. What makes this particularly dangerous is that these AI-generated outputs often have the veneer of legitimate research results—they look professional and sound plausible. However, closer inspection reveals significant gaps, inaccuracies, and missed insights that would mislead stakeholders and result in poor design decisions. The appearance of competence masks fundamental limitations that make these tools unreliable for serious research work. While we’ve found several places in the research process that can benefit from LLM usage, analysis and synthesis consistently falls short. In this talk, I can share the specific research we’re doing and explain what actually works.

Key Insights

  • AI tools frequently produce insight-shaped outputs but often lack the rigor and accuracy of trained human researchers.

  • AI moderators cannot currently assess user behavior beyond spoken words, missing key usability observations like failed or inefficient tasks.

  • Contextual elements such as environmental interruptions are critical in research but are invisible to AI tools.

  • Synthetic users generated by AI tend to produce overly positive, unrealistic feedback that can mislead product teams.

  • AI excels at finding semantic connections and grouping codes in large, already coded qualitative datasets quickly.

  • Meta-analysis of large repositories using AI can uncover recurring user themes, like change aversion, much faster than manual methods.

  • Integrating AI with organizational systems to pull in diverse data sources improves context but requires expert setup and is not yet simple.

  • AI’s context window limitations cause it to forget earlier input, affecting the accuracy of multi-turn interactions.

  • Even trained researchers must use AI outputs cautiously, vetting insights to maintain research quality.

  • Effective user research depends on human synthesis, collaboration, and contextual understanding, areas where AI currently fails.

Notable Quotes

"AI can generate insights, but it does not do them as well as a moderately trained human researcher."

"There is a world of difference between what a participant says and what they actually do, and AI misses that completely."

"AI tells you what you want to hear, which is dangerous if you’re making product decisions based on synthetic feedback."

"Our job as researchers is not making reports or interviewing users; it’s providing actionable, correct insights."

"AI tools are incentivized to produce final deliverables, but that’s an output, not the essence of research."

"AI is pretty good at finding semantic patterns among codes after human researchers have done the initial coding."

"Nobody is going to be satisfied by insight-shaped answers or high-level summaries masquerading as breakthroughs."

"AI cannot notice body language, tone, or environmental context during a research session."

"Using AI to scan large archives of research is a game changer for meta-analyses, even if it’s imperfect."

"Well-set-up AI systems pulling data from multiple company sources will have more context, but it’s still limited compared to human understanding."

Ask the Rosenbot
Jennifer Kanyamibwa
Creating the Blueprint: Growing and Building Design Teams
2018 • DesignOps Summit 2018
Gold
Sarah Auslander
Insights Panel
2022 • Civic Design 2022
Gold
Ben Davies
Expert Panel: The Principles of Research Repository Design
2022 • Advancing Research 2022
Gold
Joanna Vodopivec
One Research Team for All - Influence Without Authority
2022 • Advancing Research 2022
Gold
Phil Gilbert
A Consistent Culture of Design
2015 • Enterprise UX 2015
Gold
Frances Yllana
D.E.A.R.R. Diaries (Discipline, Experience, Architecture, Reflection + Revolution)
2022 • Civic Design 2022
Gold
Jemma Ahmed
Theme Three Intro
2023 • Advancing Research 2023
Gold
Sam Proulx
To Boldly Go: The New Frontiers of Accessibility
2022 • DesignOps Summit 2022
Gold
Jon Fukuda
All the Ops: Successful cross-functional collaboration
2025 • DesignOps Summit 2025
Gold
Courtney Kaplan
Taking it to the next level: Career paths in DesignOps
2018 • DesignOps Summit 2018
Gold
Fisayo Osilaja
[Demo] The AI edge: From researcher to strategist
2024 • Designing with AI 2024
Gold
Daniel Orbach
Zero to One: Co-Creating Operating Models with your Team
2024 • DesignOps Summit 2024
Gold
Jessamyn Edwards
Surviving Your UX Career in Enterprise Design
2021 • Enterprise Community
Bob Baxley
Theme 4: Discussion
2024 • Enterprise Experience 2020
Gold
JJ Kercher
A Roadmap for Maturing Design in the Enterprise
2018 • Enterprise Experience 2018
Gold
Kristin Skinner
Opening Keynote: Org Design for Design Orgs
2017 • DesignOps Summit 2017
Gold

More Videos

Abby Covert

"With collaborative tools, feedback has become a democratic process, and people feel more ownership of the project."

Abby Covert Tomer Sharon

Panel: Collaboration Tools

November 6, 2017

Clara Kliman-Silver

"It can be really hard to tell what is AI and what isn’t, and your average user may not know the difference."

Clara Kliman-Silver

UX Futures: The Role of Artificial Intelligence in Design

June 7, 2023

B. Pagels-Minor

"I don’t want to make those decisions anymore. I can’t compartmentalize who I am to be successful at work."

B. Pagels-Minor

Breaking the Tension: The Power of Enabling Your Employees to Show Up Authentically

June 10, 2022

Noel Lamb

"Legal impacts just about every aspect of research Ops, from GDPR to incentives to internal documentation."

Noel Lamb

Cultivating Business Partnerships to Grow Research Ops

March 21, 2022

Billy Carlson

"Throw away the rule book in early ideation phases to get as many ideas out as fast as possible."

Billy Carlson

Ideation tips for Product Managers

December 6, 2022

Tricia Wang

"Every culture has things to celebrate and improve upon; we should learn from others instead of imposing our own standards."

Tricia Wang

SCALE: Discussion

June 15, 2018

Nathan Shedroff

"Kids love to play. They challenge the world through games and simulation, so the things we create need to support play and experimentation."

Nathan Shedroff Hugh Dubberly Thomas J. McLeish

How Will Design be Taught When the Schools Shut Down?

May 8, 2026

Matt Stone

"We have to think beyond projects to the larger ecosystem and understand the ladder of intended outcomes."

Matt Stone

Scaling Empathy, A Case Study in Change Management

June 11, 2021

Uday Gajendar

"Cornelius Rushrew will set the stage and tone for how to tackle problems at scale."

Uday Gajendar

Theme 1: Introduction

June 9, 2021