Rosenverse
Human vs. machine: Testing AI’s ability to synthesize and analyze research

This video is only accessible to Gold members. Log in or register for a free Gold Trial Account to watch.

Log in Register

Most conference talks are accessible to Gold members, while community videos are generally available to all logged-in members.

Human vs. machine: Testing AI’s ability to synthesize and analyze research

Gold
Wednesday, March 11, 2026 • Advancing Research 2026

This video is featured in the AI Trends in User Research playlist.

Share the love for this talk
Human vs. machine: Testing AI’s ability to synthesize and analyze research
Speakers: Laura Klein
Link:

Summary

Nielsen Norman Group (NNG) has conducted and continues to conduct extensive research testing various large language model (LLM) tools designed for research synthesis and analysis. Our goal was to determine whether these AI-powered tools could meaningfully accelerate the work of experienced UX researchers. Through rigorous testing across multiple models and specialized research tools, we’ve found that while a few tools provide modest speed improvements for experienced researchers, none come close to replacing human expertise in research synthesis and analysis. The core problem is that these tools consistently exhibit critical flaws: they hallucinate findings, fail to identify meaningful patterns in qualitative data, cannot adequately consider nuanced research questions, and produce only superficial, high-level summaries of participant behavior. What makes this particularly dangerous is that these AI-generated outputs often have the veneer of legitimate research results—they look professional and sound plausible. However, closer inspection reveals significant gaps, inaccuracies, and missed insights that would mislead stakeholders and result in poor design decisions. The appearance of competence masks fundamental limitations that make these tools unreliable for serious research work. While we’ve found several places in the research process that can benefit from LLM usage, analysis and synthesis consistently falls short. In this talk, I can share the specific research we’re doing and explain what actually works.

Key Insights

  • AI tools frequently produce insight-shaped outputs but often lack the rigor and accuracy of trained human researchers.

  • AI moderators cannot currently assess user behavior beyond spoken words, missing key usability observations like failed or inefficient tasks.

  • Contextual elements such as environmental interruptions are critical in research but are invisible to AI tools.

  • Synthetic users generated by AI tend to produce overly positive, unrealistic feedback that can mislead product teams.

  • AI excels at finding semantic connections and grouping codes in large, already coded qualitative datasets quickly.

  • Meta-analysis of large repositories using AI can uncover recurring user themes, like change aversion, much faster than manual methods.

  • Integrating AI with organizational systems to pull in diverse data sources improves context but requires expert setup and is not yet simple.

  • AI’s context window limitations cause it to forget earlier input, affecting the accuracy of multi-turn interactions.

  • Even trained researchers must use AI outputs cautiously, vetting insights to maintain research quality.

  • Effective user research depends on human synthesis, collaboration, and contextual understanding, areas where AI currently fails.

Notable Quotes

"AI can generate insights, but it does not do them as well as a moderately trained human researcher."

"There is a world of difference between what a participant says and what they actually do, and AI misses that completely."

"AI tells you what you want to hear, which is dangerous if you’re making product decisions based on synthetic feedback."

"Our job as researchers is not making reports or interviewing users; it’s providing actionable, correct insights."

"AI tools are incentivized to produce final deliverables, but that’s an output, not the essence of research."

"AI is pretty good at finding semantic patterns among codes after human researchers have done the initial coding."

"Nobody is going to be satisfied by insight-shaped answers or high-level summaries masquerading as breakthroughs."

"AI cannot notice body language, tone, or environmental context during a research session."

"Using AI to scan large archives of research is a game changer for meta-analyses, even if it’s imperfect."

"Well-set-up AI systems pulling data from multiple company sources will have more context, but it’s still limited compared to human understanding."

Ask the Rosenbot
Megan Blocker
Positioning insight: Structuring teams, roles and careers for a changing research landscape
2025 • Advancing Research 2025
Gold
Jacqui Frey
Setting the Table for Dynamic Change
2019 • DesignOps Summit 2019
Gold
Asia Hoe
Partnering with Product: A Journey from Junior to Senior Design
2023 • Design in Product 2023
Gold
Bria Alexander
Opening Remarks
2024 • Advancing Research 2021
Gold
John Paul de Guzman
10k Screens Later: How We Became a Data-Driven Design Organization
2024 • DesignOps Summit 2024
Gold
Mary-Lynne Williams
Exit Interview #4: From Product Design Leadership to Sound Healing
2026 • Rosenfeld Community
Bria Alexander
Theme Two Intro
2022 • DesignOps Summit 2022
Gold
Ana Maria Montero Barrantes
The Authentic UX Talent Show
2024 • Enterprise Experience 2020
Gold
Corey Long
Hiring in DesignOps: A Critical Study on How to Hire and Get Hired
2024 • DesignOps Summit 2024
Gold
Dr. Jamika D. Burge
Advancing the Inclusion of Womxn in Research Practices
2022 • Advancing Research Community
Josh Clark
Sentient Design: Crafting Intelligent Interfaces with AI
2026 • Designing with AI 2026
Conference
Ovetta Sampson
Managing the Human Engagement Risks of AI
2025 • Designing with AI 2025
Gold
Rikki Teeters
Concept to code: Transforming ideas into functional products with Kiro
2026 • Designing with AI 2026
Conference
Sam Proulx
Prototype Reviews, People With Disabilities, and You
2021 • DesignOps Summit 2021
Gold
Peter Van Dijck
Building new AI skills: Creating outsized UX value with evals
2026 • Designing with AI 2026
Conference
Peter Van Dijck
Hands on AI #3: Claude Code for UX people
2025 • Rosenfeld Community

More Videos

Maria Rosala

"We made it mandatory to have a one-pager summary for every research project to make insights easy to manage."

Maria Rosala Shivanjali M.

Research Repositories

March 12, 2026

Dan Saffer

"Most AI fails because projects require near perfect accuracy, which AI can't reliably deliver."

Dan Saffer

Why AI projects fail (and what we can do about it)

May 14, 2025

Melissa Eggleston

"The effect you have on others is the most valuable currency there is."

Melissa Eggleston Maya Israni Florence Kasule Owen Seely Andrea Schneider

Practical People Skills for Building Trust on Teams and with Partners

December 9, 2021

"Find an unindicted co-conspirator in your company with a problem you can help solve, and the rest will follow."

Discussion

June 9, 2017

Satyam Kantamneni

"User comes before experience, experience before design, design before technology."

Satyam Kantamneni

Do You Have an Experience Vision?

March 23, 2023

Richard Buchanan

"Henry Ford had an intuition when he doubled workers’ wages because he realized they would buy cars and fuel the economy."

Richard Buchanan

Creativity and Principles in the Flourishing Enterprise

June 15, 2018

Victor M. Gonzalez

"Becoming a UX researcher is a process that will never end, not a fixed state you reach."

Victor M. Gonzalez

Practicing Learners and Learning Practitioners

March 10, 2021

Christian Rohrer

"We call it research goo: regulation, risk, inefficiencies that complicate research in banks."

Christian Rohrer

Research Operations at Scale

November 7, 2017

Michelle Morrison

"Sharing our playbook with engineering helped push design diversity thinking beyond our own team up to the C-suite level."

Michelle Morrison

Culture Design

May 21, 2020