All Stories

  1. Creating a Hybrid Rule and Neural Network Based Semantic Tagger Using Silver Standard Data: The PyMUSAS Framework for Multilingual Semantic Annotation
  2. Who uses mental health support forums, and why? Triangulating findings from surveys, interviews, and forum posts
  3. Improving on State-of-the-Art Models for Sentiment Analysis on Saudi-English Code-Switching Text
  4. Overview of the Second Workshop on Language Models for Low-Resource Languages (LoResLM 2026)
  5. Front Matter
  6. Front Matter
  7. Evaluating Peer Online Forums to Support Health: Ethical and Practical Challenges
  8. Impacts of Using Peer Online Forums in Mental Health: Realist Evaluation Using Mixed Methods
  9. Understanding the Needs of Moderators in Online Mental Health Forums: Realist Synthesis and Recommendations for Support
  10. Annotator Disagreement-Based Analysis for Developing Bias Benchmark Datasets in Resource-Restricted Settings
  11. Understanding Safety in Online Mental Health Forums: Realist Evaluation
  12. Co-design of moderator training: Integrating knowledge from forum moderators, users and researchers with the improving peer online forums (iPOF) project
  13. Who uses mental health support forums, and why? Triangulating findings from surveys, interviews, and forum posts
  14. Navigating Hypersexuality in Bipolar: Insights from a Corpus-Assisted Discourse Analysis of Reddit Posts
  15. ‘In my experience …’: The use of the word experience in peer online forums for mental health
  16. Using Natural Language Processing Methods to Build the Hypersexuality in Bipolar Reddit Corpus: Infodemiology Study of Reddit
  17. ParlaMint II: advancing comparable parliamentary corpora across Europe
  18. FreeTxt: A corpus-based bilingual free-text survey and questionnaire data analysis toolkit
  19. Using Natural Language Processing Methods to Build the Hypersexuality in Bipolar Reddit Corpus: Infodemiology Study of Reddit (Preprint)
  20. Understanding the Impacts of Online Mental Health Peer Support Forums: Realist Synthesis
  21. Lived experience at the core: A classification system for risk-taking behaviours in bipolar
  22. Improving shared decision-making about cancer treatment through design-based data-driven decision-support tools and redesigning care paths: an overview of the 4D PICTURE project
  23. Towards an Extensible Framework for Understanding Spatial Narratives
  24. How People With a Bipolar Disorder Diagnosis Talk About Personal Recovery in Peer Online Support Forums: Corpus Framework Analysis Using the POETIC Framework
  25. Posting patterns in peer online support forums and their associations with emotions and mood in bipolar disorder: Exploratory analysis
  26. Improving Peer Online Forums (iPOF): protocol for a realist evaluation of peer online mental health forums to inform practice and policy
  27. Semantic Tagging for the Urdu Language: Annotated Corpus and Multi-Target Classification Methods
  28. Cross-Lingual Text Reuse Detection at Document Level for English-Urdu Language Pair
  29. Correction: Social Media Monitoring of the COVID-19 Pandemic and Influenza Epidemic With Adaptation for Informal Language in Arabic Twitter Data: Qualitative Study
  30. A Comparative Study of Evaluation Metrics for Long-Document Financial Narrative Summarization with Transformers
  31. Semantic domains across topics, genders and languages
  32. Textual variations affect human judgements of sentiment values
  33. Natural Language Processing Methods and Bipolar Disorder: Scoping Review
  34. Assessment of non-directed computer-use behaviours in the home can indicate early cognitive impairment: A proof of principle longitudinal study
  35. The ParlaMint corpora of parliamentary proceedings
  36. UNLT: Urdu Natural Language Toolkit
  37. Natural Language Processing Methods and Bipolar Disorder: Scoping Review (Preprint)
  38. Multilingual Financial Word Embeddings for Arabic, English and French
  39. A Domain Based Approach to Semantic Lexicon Expansion
  40. Social Media Monitoring of the COVID-19 Pandemic and Influenza Epidemic With Adaptation for Informal Language in Arabic Twitter Data: Qualitative Study
  41. Problematising characteristicness
  42. Social Media Monitoring of the COVID-19 Pandemic and Influenza Epidemic With Adaptation for Informal Language in Arabic Twitter Data: Qualitative Study (Preprint)
  43. MasakhaNER: Named Entity Recognition for African Languages
  44. Understanding who uses Reddit: Profiling individuals with a self-reported bipolar disorder diagnosis
  45. MUMBO: MUlti-task Max-Value Bayesian Optimization
  46. Uncovering Environmental Change in the English Lake District: Using Computational Techniques to Trace the Presence and Documentation of Historical Flora
  47. Analysing Keyword Lists
  48. A Sense Annotated Corpus for All-Words Urdu Word Sense Disambiguation
  49. Retrieving, classifying and analysing narrative commentary in unstructured (glossy) annual reports published as PDF files
  50. Developing Multilingual Automatic Semantic Annotation Systems
  51. In search of meaning: Lessons, resources and next steps for computational analysis of financial discourse
  52. Leveraging Pre-Trained Embeddings for Welsh Taggers
  53. A word sense disambiguation corpus for Urdu
  54. CLEU - A Cross-Language English-Urdu Corpus and Benchmark for Text Reuse Experiments
  55. Known and unknown requirements in healthcare
  56. Multilingual Text Analysis
  57. Can you detect early dementia from an email? A proof of principle study of daily computer use to detect cognitive and functional decline
  58. A deeply annotated testbed for geographical text analysis
  59. A time-sensitive historical thesaurus-based semantic tagger for deep semantic annotation
  60. Exploring Deep Mapping Concepts: Crosthwaite’s Map and West’s Picturesque Stations
  61. The 2017 Annual Meeting of the International Genetic Epidemiology Society
  62. Measuring Short Text Reuse For The Urdu Language
  63. Lancaster A at SemEval-2017 Task 5: Evaluation metrics matter: predicting sentiment from financial news headlines
  64. Sampling labelled profile data for identity resolution
  65. lexiDB: A scalable corpus database management system
  66. Combining Mouse and Keyboard Events with Higher Level Desktop Actions to Detect Mild Cognitive Impairment
  67. COUNTER: corpus of Urdu news text reuse
  68. Towards Interactive Multidimensional Visualisations for Corpus Linguistics
  69. Reversing the Polarity with Emoticons
  70. Heterogeneous Narrative Content in Annual Reports Published as PDF Files: Extraction, Classification and Incremental Predictive Ability
  71. Metaphor, Popular Science, and Semantic Tagging: Distant reading with theHistorical Thesaurus of English
  72. Scaling out for extreme scale corpus data
  73. A Systematic Survey of Online Data Mining Technology Intended for Law Enforcement
  74. A computer-assisted study of the use of Violence metaphors for cancer and end of life by patients, family carers and health professionals
  75. Exploring fine-grained sentiment values in online product reviews
  76. Dementia and Social Sustainability: Challenges for Software Engineering
  77. The online use of Violence and Journey metaphors by patients with cancer, as compared with health professionals: a mixed methods study
  78. Geoparsing, GIS, and Textual Analysis: Current Developments in Spatial Humanities Research
  79. Guidelines for normalising Early Modern English corpora: Decisions and justifications
  80. Sentiment analysis tools should take account of the number of exclamation marks!!!
  81. Automatically Analyzing Large Texts in a GIS Environment: The Registrar General's Reports and Cholera in the 19th Century
  82. Dealing with heterogeneous big data when geoparsing historical corpora
  83. A Service-Indepenent Model for Linking Online User Profile Information
  84. Discovering affect-laden requirements to achieve system acceptance
  85. “i didn’t spel that wrong did i. Oops”
  86. Language Independent Evaluation of Translation Style and Consistency: Comparing Human and Machine Translations of Camus’ Novel “The Stranger”
  87. Crossing Boundaries: Using GIS in Literary Studies, History and Beyond
  88. Customising geoparsing and georeferencing for historical texts
  89. Who Am I? Analyzing Digital Personas in Cybercrime Investigations
  90. Children Online
  91. “i didn’t spel that wrong did i. Oops”
  92. The language of Islamic extremism
  93. Corpus Analysis of Key Words
  94. Safeguarding Cyborg Childhoods: Incorporating the On/Offline Behaviour of Children into Everyday Social Work Practices
  95. Differentiating Act from Ideology: Evidence from Messages For and Against Violent Extremism
  96. Experiments in 17th century English: manual versus automatic conceptual history
  97. What is middleware made of?
  98. Automatic error tagging of spelling mistakes in learner corpora
  99. Analyzing the semantic content and persuasive composition of extremist media: A case study of texts produced during the Gaza conflict
  100. Classification of Short Text Comments by Sentiment and Actionability for VoiceYourView
  101. Improving the precision of corpus methods
  102. Multiword expressions: hard going or plain sailing?
  103. From key words to key semantic domains
  104. An Exploratory Study of Information Retrieval Techniques in Domain Analysis
  105. A flexible framework to experiment with ontology learning techniques
  106. A framework for P2P application development
  107. 18. Automatic extraction of translation equivalents of phrasal and light verbs in English and Russian
  108. The Identification of Spelling Variants in English and German Historical Texts: Manual or Automatic?
  109. Corpus Tools and Methods, Today and Tomorrow: Incorporating Linguists' Manual Annotations
  110. Semantics-based composition for aspect-oriented requirements engineering
  111. A tool suite for aspect-oriented requirements engineering
  112. ASSIST
  113. Annotated web as corpus
  114. Measuring MWE compositionality using semantic annotation
  115. EA-Miner
  116. Shallow knowledge as an aid to deep understanding in early phase requirements engineering
  117. Comparing and combining a semantic tagger and a statistical tool for MWE extraction
  118. Artefacts as designed, artefacts as used: resources for uncovering activity dynamics
  119. Early-AIM: an approach for identifying aspects in requirements
  120. Natural Language Processing and Information Systems
  121. Extracting multiword expressions with a semantic tagger
  122. The REVERE Project: Experiments with the Application of Probabilistic NLP to Systems Engineering
  123. Assisting Requirements Recovery from Legacy Documents
  124. Comparing corpora using frequency profiling
  125. Comparing corpora using frequency profiling
  126. Social Differentiation in the Use of English Vocabulary
  127. A Flexible Framework To Experiment With Ontology Learning Techniques
  128. Supporting Law Enforcement in Digital Communities through Natural Language Analysis
  129. EA-Miner: Towards Automation in Aspect-Oriented Requirements Engineering
  130. P2P-4-DL: digital library over peer-to-peer