ZICQ
中 Log in / Sign up
Newsroom LLMs & Foundation Models #User Research #Sentiment Analysis #Multimodal Analysis #Generative AI #User Experience

arXiv Publishes Cross-Platform Analysis of Trust and Friction in Generative AI App Reviews

Avatar of Mr.Xu

By Mr.Xu

Published: · 4 views

中文阅读 (Chinese) English Version

Summary:arXiv has released a cross-platform study analyzing user trust and friction in generative AI applications, examining 17,012 English-language reviews from Google Play and the Apple App Store for six major GenAI apps (ChatGPT, Gemini, Microsoft Copilot, Claude, DeepSeek, and Perplexity). The research combines BERTopic topic modeling with RoBERTa sentiment classification, revealing that negative sentiment is concentrated in areas such as advertising, authentication, server reliability, and subscrip


Background and Objectives

In recent years, generative AI (GenAI) applications have seen rapid adoption in the consumer market. However, large-scale research on user-perceived quality, trust, and adoption barriers for these applications remains limited. This study aims to analyze user reviews to uncover key trust and friction points in mainstream GenAI applications.

Methodology

The research team collected 17,012 user reviews from Google Play and the Apple App Store for six major GenAI applications (ChatGPT, Gemini, Microsoft Copilot, Claude, DeepSeek, and Perplexity). The analysis employed BERTopic topic modeling and RoBERTa sentiment classification, supplemented by chi-square, Kruskal-Wallis, and multinomial logistic regression analyses to compare differences across applications. A stratified sample of 300 reviews was manually coded to validate the results.

Key Findings

  1. Negative Sentiment Hotspots: Advertising (91%), authentication (89%), server reliability (83%), and subscription pricing (73%) are the main areas where negative sentiment is concentrated.
  2. Application-Specific Sentiment Differences: Claude exhibits the highest negative sentiment (47.7%) but also has a highly enthusiastic user base, indicating significant polarization.
  3. Geopolitical and Data Privacy Concerns: A subset of DeepSeek reviews raised issues related to its Chinese origin, including geopolitical and data privacy concerns.
  4. Trust Friction Score: The study proposes a 'Trust Friction Score' to quantify and interpret application-specific trust and usability barriers.

Technical Highlights

  • Multimodal Analysis: Combines BERTopic topic modeling and RoBERTa sentiment classification for a nuanced analysis of user reviews.
  • Cross-Application Comparison: Utilizes chi-square, Kruskal-Wallis, and multinomial logistic regression to highlight significant differences across applications.
  • Trust Friction Score: Introduces a new quantitative metric to evaluate and explain trust and usability barriers in applications.

Industry Impact and Developer Recommendations

  1. User Experience Optimization: Developers should focus on improving user experience in areas such as advertising, authentication, server reliability, and subscription pricing to enhance user satisfaction and trust.
  2. Data Privacy and Security: For applications with geopolitical sensitivity, developers need to pay more attention to data privacy and security to build user trust.
  3. Sentiment Analysis and User Feedback: Use sentiment analysis to regularly monitor user feedback and adjust product strategies to meet changing user needs and expectations.

Conclusion

This study provides valuable insights into user trust and friction in generative AI applications and offers practical recommendations for developers to improve user experience and satisfaction.


Source: ArXiv NLP/LLM (cs.CL) (2026-09-18)

— END —

Tags: #User Research #Sentiment Analysis #Multimodal Analysis #Generative AI #User Experience

Community Comments

Loading live comments and annotations…