23 |
Nonparametric Bayesian Storyline Detection from Microtexts ...
|
|
|
|
BASE
|
|
Show details
|
|
24 |
A Kernel Independence Test for Geographical Language Variation ...
|
|
|
|
BASE
|
|
Show details
|
|
26 |
The Social Dynamics of Language Change in Online Networks ...
|
|
|
|
BASE
|
|
Show details
|
|
27 |
More emojis, less :) The competition for paralinguistic function in microblog writing
|
|
|
|
In: First Monday; Volume 21, Number 11 - 7 November 2016 ; 1396-0466 (2016)
|
|
BASE
|
|
Show details
|
|
28 |
Confounds and Consequences in Geotagged Twitter Data ...
|
|
|
|
Abstract:
Twitter is often used in quantitative studies that identify geographically-preferred topics, writing styles, and entities. These studies rely on either GPS coordinates attached to individual messages, or on the user-supplied location field in each profile. In this paper, we compare these data acquisition techniques and quantify the biases that they introduce; we also measure their effects on linguistic analysis and text-based geolocation. GPS-tagging and self-reported locations yield measurably different corpora, and these linguistic differences are partially attributable to differences in dataset composition by age and gender. Using a latent variable model to induce age and gender, we show how these demographic variables interact with geography to affect language use. We also show that the accuracy of text-based geolocation varies with population demographics, giving the best results for men above the age of 40. ... : final version for EMNLP 2015 ...
|
|
Keyword:
Computation and Language cs.CL; FOS Computer and information sciences
|
|
URL: https://arxiv.org/abs/1506.02275 https://dx.doi.org/10.48550/arxiv.1506.02275
|
|
BASE
|
|
Hide details
|
|
29 |
Overcoming Language Variation in Sentiment Analysis with Social Attention ...
|
|
|
|
BASE
|
|
Show details
|
|
30 |
Better Document-level Sentiment Analysis from RST Discourse Parsing ...
|
|
|
|
BASE
|
|
Show details
|
|
36 |
Multilingual Part-of-Speech Tagging: Two Unsupervised Approaches ...
|
|
|
|
BASE
|
|
Show details
|
|
37 |
POS induction with distributional and morphological information using a distance-dependent Chinese Restaurant Process
|
|
|
|
BASE
|
|
Show details
|
|
38 |
One Vector is Not Enough: Entity-Augmented Distributional Semantics for Discourse Relations ...
|
|
|
|
BASE
|
|
Show details
|
|
39 |
Entity-Augmented Distributional Semantics for Discourse Relations ...
|
|
|
|
BASE
|
|
Show details
|
|
|
|