Rosenverse
Human vs. machine: Testing AI’s ability to synthesize and analyze research

This video is only accessible to Gold members. Log in or register for a free Gold Trial Account to watch.

Log in Register

Most conference talks are accessible to Gold members, while community videos are generally available to all logged-in members.

Human vs. machine: Testing AI’s ability to synthesize and analyze research

Gold
Wednesday, March 11, 2026 • Advancing Research 2026

This video is featured in the AI Trends in User Research playlist.

Share the love for this talk
Human vs. machine: Testing AI’s ability to synthesize and analyze research
Speakers: Laura Klein
Link:

Summary

Nielsen Norman Group (NNG) has conducted and continues to conduct extensive research testing various large language model (LLM) tools designed for research synthesis and analysis. Our goal was to determine whether these AI-powered tools could meaningfully accelerate the work of experienced UX researchers. Through rigorous testing across multiple models and specialized research tools, we’ve found that while a few tools provide modest speed improvements for experienced researchers, none come close to replacing human expertise in research synthesis and analysis. The core problem is that these tools consistently exhibit critical flaws: they hallucinate findings, fail to identify meaningful patterns in qualitative data, cannot adequately consider nuanced research questions, and produce only superficial, high-level summaries of participant behavior. What makes this particularly dangerous is that these AI-generated outputs often have the veneer of legitimate research results—they look professional and sound plausible. However, closer inspection reveals significant gaps, inaccuracies, and missed insights that would mislead stakeholders and result in poor design decisions. The appearance of competence masks fundamental limitations that make these tools unreliable for serious research work. While we’ve found several places in the research process that can benefit from LLM usage, analysis and synthesis consistently falls short. In this talk, I can share the specific research we’re doing and explain what actually works.

Key Insights

  • AI tools frequently produce insight-shaped outputs but often lack the rigor and accuracy of trained human researchers.

  • AI moderators cannot currently assess user behavior beyond spoken words, missing key usability observations like failed or inefficient tasks.

  • Contextual elements such as environmental interruptions are critical in research but are invisible to AI tools.

  • Synthetic users generated by AI tend to produce overly positive, unrealistic feedback that can mislead product teams.

  • AI excels at finding semantic connections and grouping codes in large, already coded qualitative datasets quickly.

  • Meta-analysis of large repositories using AI can uncover recurring user themes, like change aversion, much faster than manual methods.

  • Integrating AI with organizational systems to pull in diverse data sources improves context but requires expert setup and is not yet simple.

  • AI’s context window limitations cause it to forget earlier input, affecting the accuracy of multi-turn interactions.

  • Even trained researchers must use AI outputs cautiously, vetting insights to maintain research quality.

  • Effective user research depends on human synthesis, collaboration, and contextual understanding, areas where AI currently fails.

Notable Quotes

"AI can generate insights, but it does not do them as well as a moderately trained human researcher."

"There is a world of difference between what a participant says and what they actually do, and AI misses that completely."

"AI tells you what you want to hear, which is dangerous if you’re making product decisions based on synthetic feedback."

"Our job as researchers is not making reports or interviewing users; it’s providing actionable, correct insights."

"AI tools are incentivized to produce final deliverables, but that’s an output, not the essence of research."

"AI is pretty good at finding semantic patterns among codes after human researchers have done the initial coding."

"Nobody is going to be satisfied by insight-shaped answers or high-level summaries masquerading as breakthroughs."

"AI cannot notice body language, tone, or environmental context during a research session."

"Using AI to scan large archives of research is a game changer for meta-analyses, even if it’s imperfect."

"Well-set-up AI systems pulling data from multiple company sources will have more context, but it’s still limited compared to human understanding."

Ask the Rosenbot
Lona Moore
Scaling Design Beyond Designers
2021 • Design at Scale 2021
Gold
Saara Kamppari-Miller
Theme Three Intro
2022 • DesignOps Summit 2022
Gold
Marc Rettig
Discussion
2015 • Enterprise UX 2015
Gold
Karen McGrane
AI for Information Architects: Are the robots coming for our jobs?
2024 • Rosenfeld Community
Maria Giudice
Remaking the Making Company: Moving from Product to Experience
2016 • Enterprise UX 2016
Gold
Aleksandra Korczynska
Survey Tools
2026 • Advancing Research 2026
Gold
Shipra Kayan
Make your research synthesis speedy and more collaborative using a canvas
2025 • Rosenfeld Community
Kate Towsey
ResearchOps AMA with Kate Towsey & Jake Burghardt
2025 • Advancing Research Community
Juhan Sonin
Design Now! The Agenda for Action
2025 • Rosenfeld Community
Uday Gajendar
Leading through the long tail of trauma
2022 • Advancing Research Community
Doug Powell
DesignOps and the Next Frontier: Leading Through Unpredictable Change
2025 • DesignOps Summit 2025
Gold
Tara Tressel
Investigating qualitative depth of AI-moderated interviews
2026 • Advancing Research 2026
Gold
Gina Mendolia
Therapists, Coaches, and Grandmas: Techniques for Service Design in Complex Systems
2024 • Advancing Service Design 2024
Gold
Leisa Reichelt
Opening Keynote: Operating in Context
2018 • DesignOps Summit 2018
Gold
Shreya Dhawan
Making service tangible: the fastest path to higher performance
2025 • Advancing Service Design 2025
Gold
Russ Unger
Onboarding: The Ecosystem, not the Afterthought
2017 • DesignOps Summit 2017
Gold

More Videos

Gabriela Barneva

"Accessibility often happens too late, as an afterthought, making updates costly and disconnected from lived experience."

Gabriela Barneva

Operationalizing Inclusive Design in Design Ops

September 11, 2025

Luz Bratcher

"Busyness dehumanizes us when we tie our worth only to what we produce."

Luz Bratcher

This Is a Talk for Tired People

June 10, 2022

Stefanie Owens

"Injecting a culture of agile iteration sets the precedent that things can and will change as the team learns."

Stefanie Owens

Optimizing for Outcomes: Transformation Design in Systems at Scale

December 4, 2024

Isaac Heyveld

"Communication is an integral part of the chief of staff role, whether with leadership, the org, or partners."

Isaac Heyveld

Expand DesignOps Leadership as a Chief of Staff

September 8, 2022

Sarah Auslander

"Incremental change was used to drive a radical shift in policy despite initial resistance."

Sarah Auslander

Incremental Steps to Drive Radical Innovation in Policy Design

November 18, 2022

Ovetta Sampson

"Minimize anthropomorphism in AI design so users know they are engaging with machines, not humans."

Ovetta Sampson

Managing the Human Engagement Risks of AI

June 10, 2025

Jennifer Fraser

"An ecosystem map depicts relationships between animate and inanimate objects in a system representing value exchanges."

Jennifer Fraser

What would Emmy Noether Do? Math, Models and Mulling in UX Research

March 29, 2023

Matt Duignan

"Researchers can publish directly without intermediation to avoid blocking knowledge sharing."

Matt Duignan

HITS, Microsoft's internal human insight system: From research library to living body of knowledge

July 16, 2019

Satyam Kantamneni

"What if you get a notification when you drive into a dealership and all your details are ready?"

Satyam Kantamneni

Do You Have an Experience Vision?

March 23, 2023