Study finds trust and friction issues in major generative AI app reviews
Researchers examined 17,012 English‑language reviews from Google Play and the Apple App Store for six leading generative AI applications—ChatGPT, Gemini, Microsoft Copilot, Claude, DeepSeek and Perplexity. Using BERTopic for topic modeling and RoBERTa for sentiment classification, they validated the approach against a stratified sample of 300 human‑coded reviews and applied chi‑square,…
Key points
- Analysis of 17,012 English reviews from Google Play and Apple App Store across six GenAI apps.
- Negative sentiment clusters in advertising (91%), authentication (89%), server reliability (83%) and subscription pricing (73%).
- Claude shows highest negative sentiment at 47.7% and strong polarization among users.
Negative sentiment was concentrated in four areas: advertising (91%), authentication (89%), server reliability (83%) and subscription pricing (73%). Sentiment varied significantly across applications; Claude recorded the highest negative sentiment at 47.7% while also showing a strongly enthusiastic user base, indicating pronounced polarization. A subset of DeepSeek reviews raised geopolitical and data‑privacy concerns tied to its Chinese origin. The authors introduced a Trust Friction Score to summarize application‑specific trust and usability barriers into interpretable dimensions, offering actionable evidence for developers and policymakers.
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us.
More in Research
All →- New framework optimizes LLM inference costs via adaptive model activation · 4 src
- Qwen3.5-4B outperforms larger LLMs on new user-side conflict benchmark · 1 src
- Neo-Classic benchmark evaluates linguistic-aesthetic reasoning in Classical Chinese poetry · 1 src
- Study finds PCA can detect stylistic axes in LLM activations without training · 1 src
- Study finds causal control in subliminal prompting varies by model depth · 1 src
Comments
via GitHub Discussions