lnu.sePublications
Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
Reuse of social media data in corpus linguistics
University of Eastern Finland, Finland;Center for Data Intensive Sciences and Applications, Sweden. (DISA)ORCID iD: 0000-0003-3123-6932
University of Eastern Finland.ORCID iD: 0000-0002-5239-8407
2025 (English)In: International Journal of Corpus Linguistics, ISSN 1384-6655, E-ISSN 1569-9811, Vol. 30, no 2, p. 171-194Article in journal (Refereed) Published
Abstract [en]

The use of very large social media datasets in corpus linguistics has obviousbenefits. Such data represent a novel source of evidence when compared with structured digital text corpora. However, there is a clear need to assess critically how the effective reuse of data can be handled, how findings can be reproduced, and how results can be generalized. A relevant question concerns the presentation of data to ensure reproducibility and replicability. This article surveys the state-of-the-art of descriptions of data collection and methodological transparency in 30 studies that used Twitter/X as their data. The empirical section investigates how easy it would be to reproduce a study based on these descriptions. While we concentrate on evidence from one social media application, the discussion continues to a presentation of concrete steps that might be used to improve data management related to the reuse, discovery, and evaluation of social media data in general.

Place, publisher, year, edition, pages
John Benjamins Publishing Company, 2025. Vol. 30, no 2, p. 171-194
Keywords [en]
social media data, replicability, reproducibility, metadata, research infrastructures
National Category
Studies of Specific Languages
Research subject
Humanities, English
Identifiers
URN: urn:nbn:se:lnu:diva-139904DOI: 10.1075/ijcl.24136.laiISI: 001507856200001Scopus ID: 2-s2.0-105009278659OAI: oai:DiVA.org:lnu-139904DiVA, id: diva2:1973349
Available from: 2025-06-19 Created: 2025-06-19 Last updated: 2025-10-13Bibliographically approved

Open Access in DiVA

No full text in DiVA

Other links

Publisher's full textScopus

Authority records

Laitinen, Mikko

Search in DiVA

By author/editor
Laitinen, MikkoRautionaho, Paula
In the same journal
International Journal of Corpus Linguistics
Studies of Specific Languages

Search outside of DiVA

GoogleGoogle Scholar

doi
urn-nbn

Altmetric score

doi
urn-nbn
Total: 174 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf