Yelp13 Sentiment Eval

This benchmark evaluates a model's ability to predict sentiment on long, complex documents by leveraging discourse structure. It probes whether incorporating hierarchical discourse trees improves sentiment classification and regression over standard sequential baselines, particularly for longer texts where sentiment is more subtle and diverse. Use when the user wants to benchmark on Yelp'13, or asks about evaluating this task. Reports accuracy.

qhjqhj00 616f505 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/yelp13-sentiment-eval commit 616f505207

Frequently asked questions

npx skillmds add qhjqhj00/yelp13-sentiment-eval