Rosenverse
Human vs. machine: Testing AI’s ability to synthesize and analyze research

This video is only accessible to Gold members. Log in or register for a free Gold Trial Account to watch.

Log in Register

Most conference talks are accessible to Gold members, while community videos are generally available to all logged-in members.

Human vs. machine: Testing AI’s ability to synthesize and analyze research

Gold
Wednesday, March 11, 2026 • Advancing Research 2026
Share the love for this talk
Human vs. machine: Testing AI’s ability to synthesize and analyze research
Speakers: Laura Klein
Link:

Summary

Nielsen Norman Group (NNG) has conducted and continues to conduct extensive research testing various large language model (LLM) tools designed for research synthesis and analysis. Our goal was to determine whether these AI-powered tools could meaningfully accelerate the work of experienced UX researchers. Through rigorous testing across multiple models and specialized research tools, we’ve found that while a few tools provide modest speed improvements for experienced researchers, none come close to replacing human expertise in research synthesis and analysis. The core problem is that these tools consistently exhibit critical flaws: they hallucinate findings, fail to identify meaningful patterns in qualitative data, cannot adequately consider nuanced research questions, and produce only superficial, high-level summaries of participant behavior. What makes this particularly dangerous is that these AI-generated outputs often have the veneer of legitimate research results—they look professional and sound plausible. However, closer inspection reveals significant gaps, inaccuracies, and missed insights that would mislead stakeholders and result in poor design decisions. The appearance of competence masks fundamental limitations that make these tools unreliable for serious research work. While we’ve found several places in the research process that can benefit from LLM usage, analysis and synthesis consistently falls short. In this talk, I can share the specific research we’re doing and explain what actually works.

Key Insights

  • AI tools frequently produce insight-shaped outputs but often lack the rigor and accuracy of trained human researchers.

  • AI moderators cannot currently assess user behavior beyond spoken words, missing key usability observations like failed or inefficient tasks.

  • Contextual elements such as environmental interruptions are critical in research but are invisible to AI tools.

  • Synthetic users generated by AI tend to produce overly positive, unrealistic feedback that can mislead product teams.

  • AI excels at finding semantic connections and grouping codes in large, already coded qualitative datasets quickly.

  • Meta-analysis of large repositories using AI can uncover recurring user themes, like change aversion, much faster than manual methods.

  • Integrating AI with organizational systems to pull in diverse data sources improves context but requires expert setup and is not yet simple.

  • AI’s context window limitations cause it to forget earlier input, affecting the accuracy of multi-turn interactions.

  • Even trained researchers must use AI outputs cautiously, vetting insights to maintain research quality.

  • Effective user research depends on human synthesis, collaboration, and contextual understanding, areas where AI currently fails.

Notable Quotes

"AI can generate insights, but it does not do them as well as a moderately trained human researcher."

"There is a world of difference between what a participant says and what they actually do, and AI misses that completely."

"AI tells you what you want to hear, which is dangerous if you’re making product decisions based on synthetic feedback."

"Our job as researchers is not making reports or interviewing users; it’s providing actionable, correct insights."

"AI tools are incentivized to produce final deliverables, but that’s an output, not the essence of research."

"AI is pretty good at finding semantic patterns among codes after human researchers have done the initial coding."

"Nobody is going to be satisfied by insight-shaped answers or high-level summaries masquerading as breakthroughs."

"AI cannot notice body language, tone, or environmental context during a research session."

"Using AI to scan large archives of research is a game changer for meta-analyses, even if it’s imperfect."

"Well-set-up AI systems pulling data from multiple company sources will have more context, but it’s still limited compared to human understanding."

Ask the Rosenbot
Erika Flowers
AI-Readiness: Preparing NASA for a Data-Driven, Agile Future
2025 • Designing with AI 2025
Gold
Ryan Matthew
Bridging Design and Code: AI-Powered Design System Integration
2025 • Rosenfeld Community
Jane Davis
Strategic Shifts and Innovations in User Research: Navigating Challenges and Opportunities
2025 • Advancing Research 2025
Gold
Tristin Oldani
Turning awareness into action with Climate UX
2025 • Climate UX Interest Group
Elizabeth Churchill
Exploring Cadence: You, Your Team, and Your Enterprise
2017 • Enterprise Experience 2017
Gold
Kim Lenox
Leading Distributed Global Teams
2019 • Enterprise Community
Bob Baxley
Theme 4: Intro
2024 • Enterprise Experience 2020
Gold
Sarah Kinkade
Design Management Models in the Face of Transformation
2022 • Design at Scale 2022
Gold
Gregg Bernstein
Opportunistic Research with Gregg Bernstein
2019 • Advancing Research Community
Mike Brzozowski
UX in everyday products: Empowering climate conscious choices
2024 • Climate UX Interest Group
Ed Mullen
Designing the Unseen: Enabling Institutions to Build Public Trust
2022 • Civic Design 2022
Gold
Dr. Jamika D. Burge
The Future of Research: Bridging the Gaps
2021 • Advancing Research Community
Victor Lombardi
Bridging Design and Climate Science
2024 • Climate UX Interest Group
Jackie Velasquez-Ross
Talent Acquisition and Our Responsibility
2020 • DesignOps Community
Sam Proulx
Understanding Screen Readers on Mobile: How And Why to Learn from Native Users
2023 • DesignOps Summit 2023
Gold
Maverick Chan
From Doodle to Demo: AI as Our Storytelling Partner
2025 • Rosenfeld Community

More Videos

Husani Oakley

"In the enterprise, complexity and chaos are just another Tuesday."

Husani Oakley

Theme Two Intro

June 6, 2023

Mansi Gupta

"In that moment, I looked down at my chest and failed to find a breast pocket of my own."

Mansi Gupta

Drawing from Feminist Practice to Make Inclusive Design Operational

September 9, 2022

Sam Proulx

"Focus changes are critical to inform screen reader users when dialogs open or content changes."

Sam Proulx

Designing For Screen Readers: Understanding the Mental Models and Techniques of Real Users

December 10, 2021

Sarah Coyle

"We are just starting to do regular and frequent reporting across the entire design organization."

Sarah Coyle

Design and Analytics with Sarah Coyle

July 30, 2020

Alan Williams

"Categorical eligibility allows us to communicate benefit access without asking for personal data directly."

Alan Williams Rose Deeb

Designing essential financial services for those in need

February 10, 2022

Dan Donald

"The value of a design system looks different depending on if you’re a developer, designer, or stakeholder."

Dan Donald

Design Systems as a Vehicle for Systemic Change

June 1, 2023

Samuel Martin

"The same reason they hire us is the very reason why they fire us, because when things get real, it sometimes gets hard."

Samuel Martin

Co-Design vs Faux-Design: Navigating the Complexities of Sharing Power in Co-Design

March 27, 2026

Phil Hesketh

"Language in consent forms doesn’t have to be complicated; plain language and user-centered design can make it accessible."

Phil Hesketh

Designing Accessible Research Workflows

September 29, 2021

David Conrad

"Designers should act like therapists or journalists when negotiating political divides around data governance."

David Conrad

The Feeling of Data

September 14, 2023