Code-switching in Irish tweets: A preliminary analysis
Explore this paper's citation graph
Summary
The annotation of (English) code- Switching in a corpus of 1496 Irish tweets is reported on and a computational analysis of the nature of code-switching amongst Irish speaking Twitter users is provided, with a view to providing a basis for future linguistic and socio-linguistic studies.
- Type
- article
- Published
- 2019-08-01
- Cited by
- 13
- References
- 34
- Access
- Open access
- OpenAlex
- https://openalex.org/W3003326963
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:202605494
Keywords
Code-switching, Irish, Computer science, Annotation, Code (set theory)
References
- Subsegmental language detection in Celtic language text
- A tagging algorithm for mixed language identification in a noisy domain
- Histories and Pseudo-Histories of the Insular Middle Ages
- Code-Mixing in Biliterate and Multiliterate Irish Literary Texts
- Bilingual: Life and Reality
- Learning Morphology with Morfette
- A social network analysis of Irish language use in social media
- Computational Sociolinguistics: A Survey
- Processing of Sentences With Intra-Sentential Code-Switching
- Codeswitching, identity and ownership in Irish radio comedy
- Code‐switching and borrowing in Irish1
- A Coefficient of Agreement for Nominal Scales
- Survey Article: Inter-Coder Agreement for Computational Linguistics
- Part-of-Speech Tagging for English-Spanish Code-Switched Text
- Squibs and Discussions: The Kappa Statistic: A Second Look
- Part-of-Speech Tagging for Twitter: Annotation, Features, and Experiments
- The measurement of observer agreement for categorical data.
- Overview for the Second Shared Task on Language Identification in Code-Switched Data
- Minority Language Twitter: Part-of-Speech Tagging and Analysis of Irish Tweets
- Challenges of Computational Processing of Code-Switching
Cited by
- Can Multilingual Language Models Transfer to an Unseen Dialect? A Case Study on North African Arabizi
- Treebanking User-Generated Content: A Proposal for a Unified Representation in Universal Dependencies
- Building a User-Generated Content North-African Arabizi Treebank: Tackling Hell
- Treebanking user-generated content: a UD based overview of guidelines, corpora and unified recommendations
- Harnessing Indigenous Tweets: The Reo Māori Twitter corpus
- TwittIrish: A Universal Dependencies Treebank of Tweets in Modern Irish
- Across the Cyberwaves: Twitter Campaigns for Gaeilge
- Code-switching functions in online advertisements on Snapchat
- Lexical tonal effects in code-switching: A comparative study of Cantonese, Mandarin, and Vietnamese switching with English
- Representativeness as a Forgotten Lesson for Multilingual and Code-switched Data Collection and Preparation
- Synthetic Fluency: Hallucinations, Confabulations, and the Creation of Irish Words in LLM-Generated Translations
- ASCEND: A Spontaneous Chinese-English Dataset for Code-switching in Multi-turn Conversation
- Representativeness as a Forgotten Lesson for Multilingual and Code-switched Data Collection and Preparation
- An Online Linguistic Analyser for Scottish Gaelic
Related papers
- Irish on the World Wide Web: Searches and sites
- Review on Sentiment Analysis of Indian Languages with a Special Focus on Code Mixed Indian Languages
- Identifying Languages at the Word Level in Code-Mixed Indian Social Media Text
- XRCE Personal Language Analytics Engine for Multilingual Author Profiling
- Tweeting and Being Ironic in the Debate about a Political Reform: the French Annotated Corpus TWitter-MariagePourTous
- Sentiment Analysis of Code-Mixed Roman Urdu-English Social Media Text using Deep Learning Approaches
- Bilingual Multimedia: Some Challenges for Teachers