
The UGSC-ML dataset is the multilingual version of the User Gold Standard Corpus (UGSC)for sustainable urban mobility sentiment analysis. It consists of 375 English transport-related user reviews from TripAdvisor withsentence-aligned translations in Spanish, French, German, and Italian, yielding1,875 instances across five languages. The CardiffNLP XLM-RoBERTa model (`cardiffnlp/twitter-xlm-roberta-base-sentiment`)is applied in zero-shot domain transfer mode: pre-trained on CC100 and fine-tunedon Twitter sentiment data, then applied without any task- or domain-specific fine-tuningto transport reviews.
sustainable urban mobility, UGSC, sentiment analysis, XLM-RoBERTa, zero-shot domain transfer, cross-lingual evaluation, multilingual NLP, user-generated content, confidence-based assessment
Twitter Data
sustainable urban mobility, UGSC, sentiment analysis, XLM-RoBERTa, zero-shot domain transfer, cross-lingual evaluation, multilingual NLP, user-generated content, confidence-based assessment
Twitter Data
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
