4 AI models · 20 ranked · updated aug 10, 2026
Best UX Research Tools, according to AI (2026).
Maze is the answer at the latest close: the consensus #1 at a score of 75 across all 4 models.
This guide is built from the recorded answers of ChatGPT, Claude, Gemini, Perplexity to the real questions buyers ask about ux research tools. We logged 560 of them this period, covering 20 tools. We report what the models said, in the order they said it. Nobody paid to be here, and we don't add opinions of our own.
Prefer the raw board? See the full model-by-model ranking. Every rank, every score, every close.
The ranked list
Maze holds the consensus #1 at a score of 75, with Claude and Gemini and Perplexity placing it first outright.
It climbed 1 position at the latest close.
Maze — My strongest general recommendation for most product teams because it combines rapid prototype, website, mobile, survey, card-sort, tree-test, moderated, and unmoderated research in one comparatively approachable workflow.
ChatGPT, this closeMaze – The strongest all-around pick for most product teams; it's an AI-first, all-in-one platform covering prototype testing, surveys, card sorting, tree testing, and AI-moderated interviews, so it replaces several point tools at once.
Claude, this closeUserTesting ranks #2 on consensus with a score of 61, scoring best on ChatGPT (#1).
It slipped 1 position at the latest close.
UserTesting — Choose this when high-volume participant access, mature enterprise capability, and both live and self-guided testing matter more than budget simplicity.
ChatGPT, this closeUserTesting – The enterprise market leader for moderated, video-based research, best if you need deep "why" insights at scale with robust panel access and reporting.
Claude, this closeDovetail ranks #3 on consensus with a score of 43, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #9 on Perplexity, a spread of 8.
It entered the ranking at the latest close. A new name in the models' answers.
Dovetail — The best research repository and synthesis layer for teams that already collect data in Zoom, UserTesting, surveys, or elsewhere and need reusable, evidence-linked organizational memory.
ChatGPT, this closeDovetail – The leading research repository/synthesis tool, worth adding once you're running enough studies that insights need centralized tagging and cross-study search.
Claude, this closeLookback ranks #4 on consensus with a score of 34, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #7 on Claude, a spread of 6.
It entered the ranking at the latest close. A new name in the models' answers.
Lookback — Best for teams that prioritize rich qualitative sessions and want asynchronous or AI-moderated testing to retain more of the participant’s “why,” not just completion metrics.
ChatGPT, this closeLookback – A solid specialist for live, moderated remote interviews and usability sessions with strong recording/transcription features.
Claude, this closeHotjar ranks #5 on consensus with a score of 31, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
It is the most contested name in this category. It carries the widest cross-model disagreement on the board.
Hotjar – The go-to for behavioral analytics (heatmaps, session recordings, on-site surveys), essential if your gap is understanding in-product behavior rather than running formal studies.
Claude, this closeHotjar: It is the go-to choice for gathering passive behavioral insights via heatmaps and session recordings.
Gemini, this closeLyssna ranks #6 on consensus with a score of 19, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Claude, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Lyssna — A very good value-oriented, broad testing platform for small-to-mid-sized teams, especially those doing Figma validation plus card sorting and tree testing.
ChatGPT, this closeLyssna (formerly UsabilityHub) – A lightweight, budget-friendly choice for fast unmoderated tests like five-second tests, first-click testing, and card sorting, best for smaller teams needing quick directional feedback.
Claude, this closeQualtrics XM Strategy & Research ranks #7 on consensus with a score of 18, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Gemini, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Qualtrics XM Strategy & Research — Best for teams that want an all-in-one research suite with strong survey and research-program capabilities, and it is rated 4.4/5 on G2 in the cited source.
Perplexity, this closePPLX#3compare head-to-headSprig ranks #8 on consensus with a score of 15, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Sprig — The strongest option for continuous in-product research, since it can target surveys and replays based on actual user events and connect feedback to the behavior immediately around it.
ChatGPT, this closeSprig: Sprig excels at continuous discovery through targeted, in-product micro-surveys and AI-powered analysis.
Gemini, this closeGemini#4compare head-to-headUXArmy ranks #9 on consensus with a score of 15, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Gemini, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
UXArmy — Best for end-to-end usability testing plus AI-assisted insights, with strong aggregate ratings and features like task-based testing, video feedback, and AI summaries.
Perplexity, this closePPLX#4compare head-to-headUser Interviews ranks #10 on consensus with a score of 13, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
User Interviews — The best specialist purchase for research operations and participant recruitment, particularly if your actual customers and a governed internal panel are strategically important.
ChatGPT, this closeUser Interviews – Dominant for participant recruiting and panel management, a smart add-on if recruiting quality participants quickly is your bottleneck.
Claude, this closeClaude#5compare head-to-headLyssna (formerly UsabilityHub) ranks #11 on consensus with a score of 11, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Lyssna (formerly UsabilityHub) – A lightweight, budget-friendly choice for fast unmoderated tests like five-second tests, first-click testing, and card sorting, best for smaller teams needing quick directional feedback.
Claude, this closeClaude#6compare head-to-headPlaybookUX ranks #12 on consensus with a score of 11, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Gemini, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
PlaybookUX — Best for flexible qualitative interviews and surveys, especially for teams that want an accessible all-purpose research tool without enterprise complexity.
Perplexity, this closePPLX#6compare head-to-headdscout ranks #13 on consensus with a score of 8, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
dscout — Best for deep mobile, diary, and longitudinal research where in-the-moment photos, videos, repeated activities, and real-world context are more valuable than quick prototype metrics.
ChatGPT, this closedscout – The niche leader for diary studies and mobile ethnography, useful if you need longitudinal, in-context qualitative data.
Claude, this closeClaude#8compare head-to-headOptimal Workshop ranks #14 on consensus with a score of 8, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Optimal Workshop — The strongest specialist tool for information architecture work—especially serious card sorting and tree testing—rather than a full replacement for a general UX-research platform.
ChatGPT, this closeOptimal Workshop: It remains the unmatched industry standard for information architecture research like card sorting and tree testing.
Gemini, this closeGemini#8compare head-to-headTypeform ranks #15 on consensus with a score of 8, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Gemini, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Typeform — Best for surveys and simple feedback collection, especially when you need a polished respondent experience more than a full research platform.
Perplexity, this closePPLX#8compare head-to-headCondens ranks #16 on consensus with a score of 7, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Condens: This is a highly collaborative, user-friendly qualitative data analysis tool that serves as a great lightweight alternative for research repositories.
Gemini, this closeGemini#9compare head-to-headOptimal ranks #17 on consensus with a score of 7, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Optimal Workshop — The strongest specialist tool for information architecture work—especially serious card sorting and tree testing—rather than a full replacement for a general UX-research platform.
ChatGPT, this closeOptimal – The specialist choice if information architecture work (card sorting, tree testing) makes up the bulk of your research; skip it if IA is only occasional.
Claude, this closeClaude#9compare head-to-headAmplitude ranks #18 on consensus with a score of 6, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Gemini, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Amplitude — Best for product analytics when you want to pair UX research with behavioral data, but it is less of a pure research tool than the options above.
Perplexity, this closePPLX#10compare head-to-headGreat Question ranks #19 on consensus with a score of 6, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Great Question — The best consolidation choice for teams that need research CRM, recruiting, scheduling, incentives, study execution, and an insight repository rather than a collection of separate tools.
ChatGPT, this closeGreat Question: It streamlines research operations by seamlessly handling participant recruitment, scheduling, and incentive payouts.
Gemini, this closeGemini#10compare head-to-headQualtrics XM (User Experience Research) ranks #20 on consensus with a score of 6, scoring best on ChatGPT (#1).
The models disagree about it more than most: #1 on ChatGPT but #11 on Perplexity, a spread of 10.
It entered the ranking at the latest close. A new name in the models' answers.
Qualtrics XM (User Experience Research) – The enterprise-grade, quantitative-forward option that ties research into a broader customer experience/analytics platform, best for large orgs already in the Qualtrics ecosystem.
Claude, this closeClaude#10compare head-to-head
Questions people ask.
What is the best ux research tools according to AI?
Maze holds the consensus #1 at the latest close with a score of 75 out of 100, ahead of UserTesting. The models don't fully agree: 1 different brand is crowned #1 across the four models.
Which ux research tool does ChatGPT recommend first?
ChatGPT's current #1 for ux research tools is shown in the model column of the full ranking.
Which ux research tool does Claude recommend first?
Claude currently places Maze at #1 for ux research tools, per the latest close's recorded answers.
How are these rankings measured?
We ask each model the same buying questions on every run (560 queries this period), record the full answers, and score each named brand 0-100 by how early and how consistently it appears. Every question runs through the official model APIs, with web search on, not through the consumer chat apps, so nothing is personalized to a user. Each model is scored independently; the consensus blends all 4.
Do the AI models agree with each other?
At the top, yes. Every model crowns the same #1 at the latest close. Further down the board they diverge, which is why each brand carries a spread figure: the gap between its best and worst model rank.
This is the guide. The record has more.
The full ranking shows every model’s column side by side, 12 weeks of movement, and the methodology behind every number.