REVIEW 2 major objections 2 minor 300 references
NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models
T0 review · 2 major / 2 minor · reviewed 2026-06-28 · grok-4.3
Pith's one-line read Vision-language models match experts on coarse human motion ratings but fail on fine-grained part-level judgments.
desk verdict NextMotionQA adds useful structure to motion benchmarks but the fine-grained VLM judge limits rest on expert labels whose reliability is not shown. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The NextMotionQA benchmark with its three tasks (multiple-choice QA, video captioning, fine-grained error correction) stratified by semantic axes and complexity levels, plus the use of Cohen's κ to measure VLM-expert agreement on judging tasks.
What would settle it
Re-annotating a subset of the benchmark videos by an independent group of experts and finding substantially different agreement rates between VLMs and those new annotations on the fine-grained tasks.
Extended reading notes
Core claim
NextMotionQA is a benchmark built with a semi-automated expert-verified process that includes three complementary tasks structured across three core semantic axes and stratified into three complexity levels. Extensive tests on twelve VLMs uncover capability gaps invisible under single-task evaluations. VLMs align strongly with expert ratings on coarse criteria with Cohen's κ equal to 0.70 but break down on fine-grained part-level judgment with κ equal to 0.10.
Load-bearing premise
The semi-automated expert-verified dataset creation produces annotations without systematic biases or ambiguities that would distort the measured gaps in model performance.
Editorial extensions
If this is right
- Single-task evaluations hide real weaknesses in VLMs for human motion understanding.
- VLMs can serve as reliable judges only for coarse motion criteria, not detailed analysis.
- Benchmarks with explicit complexity levels and multiple tasks are required to diagnose VLM limits accurately.
Reading between the lines
- Applications in robotics and animation that rely on fine motion details may still need human oversight even when using VLMs.
- Future model training could target the specific failure modes identified in part-level motion judgment.
- The benchmark structure could be adapted to test motion understanding in non-human domains such as animals or objects.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper introduces NextMotionQA, a benchmark for human motion understanding in VLMs comprising three tasks (multiple-choice QA, video captioning, fine-grained error correction) organized along three semantic axes and three complexity levels. It uses a semi-automated expert-verified pipeline to create the dataset and evaluates twelve VLMs, reporting that they align well with expert ratings on coarse criteria (Cohen's κ=0.70) but degrade sharply on fine-grained part-level judgments (κ=0.10). The work positions this as both a diagnostic benchmark and a test of VLMs as judges for text-to-motion evaluation.
Significance. If the expert annotations prove reliable, the multi-task, multi-axis, multi-level design supplies a finer-grained diagnostic than prior motion benchmarks, exposing specific VLM failure modes invisible in single-task evaluations and clarifying the regime where VLM judges remain trustworthy. The semi-automated creation process itself is a practical engineering contribution for scalable annotation.
major comments (2)
- [Dataset creation / annotation protocol] Dataset creation section: the semi-automated expert-verified pipeline is described as producing unambiguous annotations, yet no inter-annotator agreement metrics (Cohen's κ, Fleiss' κ, or equivalent) are reported specifically for the fine-grained error-correction task or part-level judgments. This is load-bearing for the central claim, because the reported drop from κ=0.70 (coarse) to κ=0.10 (fine) cannot be attributed to VLM limitations unless expert labels are shown to be a stable gold standard rather than noisy or ambiguous.
- [VLM judge evaluation subsection] Results on VLM-as-judge (the κ comparison): the paper states the coarse/fine-grained κ values but does not report the number of expert annotators, number of items rated, or exact rating protocol used for the fine-grained condition. Without these, the magnitude of the degradation cannot be assessed for statistical robustness or potential confounds such as differing item difficulty distributions.
minor comments (2)
- [Introduction / benchmark overview] The three semantic axes and three complexity levels are introduced in the abstract and overview but would benefit from an explicit table or diagram early in the paper showing how tasks map onto axes × levels.
- [Task definitions] Notation for the three tasks (MCQA, captioning, error correction) is used consistently but the exact prompt templates or output formats for each are not reproduced in a single reference table, complicating replication.
Simulated Author's Rebuttal
We thank the referee for the constructive feedback on the annotation protocol and evaluation details. We address each major comment below.
read point-by-point responses
-
Referee: [Dataset creation / annotation protocol] Dataset creation section: the semi-automated expert-verified pipeline is described as producing unambiguous annotations, yet no inter-annotator agreement metrics (Cohen's κ, Fleiss' κ, or equivalent) are reported specifically for the fine-grained error-correction task or part-level judgments. This is load-bearing for the central claim, because the reported drop from κ=0.70 (coarse) to κ=0.10 (fine) cannot be attributed to VLM limitations unless expert labels are shown to be a stable gold standard rather than noisy or ambiguous.
Authors: We agree that explicit inter-annotator agreement metrics are necessary to substantiate the reliability of the expert labels as a gold standard. The current manuscript does not report these metrics for the fine-grained tasks. In the revision we will add Cohen's κ values computed among the expert annotators for both coarse and fine-grained conditions, confirming consistency of the labels and supporting attribution of the VLM degradation to model limitations. revision: yes
-
Referee: [VLM judge evaluation subsection] Results on VLM-as-judge (the κ comparison): the paper states the coarse/fine-grained κ values but does not report the number of expert annotators, number of items rated, or exact rating protocol used for the fine-grained condition. Without these, the magnitude of the degradation cannot be assessed for statistical robustness or potential confounds such as differing item difficulty distributions.
Authors: We acknowledge the omission of these procedural details. The revised manuscript will specify the number of expert annotators, the number of items rated under the fine-grained protocol, and the exact rating instructions and scale used. These additions will enable readers to evaluate statistical robustness and rule out confounds. revision: yes
Circularity Check
No circularity: empirical benchmark with direct expert comparisons
full rationale
The paper introduces NextMotionQA as a new benchmark with three tasks (MCQA, captioning, error correction) stratified by semantic axes and complexity, then reports empirical VLM evaluations and Cohen's κ alignments with expert ratings (0.70 coarse, 0.10 fine-grained). No equations, fitted parameters, predictions derived from inputs, or self-citation chains appear in the provided text. The central claims rest on dataset construction and direct measurement against external expert annotations rather than any reduction to self-referential definitions or renamings. The absence of mathematical derivations or load-bearing self-citations makes the work self-contained against external benchmarks.
Assumptions & free parameters
assumptions (2)
- domain assumption Expert verification ensures the quality and lack of ambiguity in the dataset annotations.
- domain assumption The three semantic axes and three complexity levels provide a comprehensive coverage of human motion understanding.
Cite this review
Pith. "Pith review of NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models." pith.science (2026). https://pith.science/paper/MHKVTQM5
@misc{pith2026260604773,
author = {Pith},
title = {Pith review of: NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models},
year = {2026},
howpublished = {\url{https://pith.science/paper/MHKVTQM5}},
note = {Machine review of arXiv:2606.04773}
}
read the original abstract
Reliable evaluation of human motion understanding is fundamental to advancing embodied AI, robotics, and animation. However, existing benchmarks suffer from coarse semantic granularity, undifferentiated difficulty, limited annotation quality, and pervasive answer ambiguity, leaving them unable to diagnose where current models fail. To bridge this gap, we introduce NextMotionQA, a comprehensive benchmark that leverages vision-language models (VLMs) for semi-automated, expert-verified dataset. NextMotionQA features three complementary tasks: multiple-choice question answering, video captioning, and fine-grained error correction. Each task is systematically structured across three core semantic axes and stratified into three task complexity levels. Our extensive evaluation of twelve representative VLMs uncovers critical capability gaps and weakness that remain invisible under conventional, single-task evaluations. In a complementary direction, recent work has begun using VLMs as judges for text-to-motion evaluation; we ask whether they show the same degradation under harder tasks. We find that VLMs align strongly with expert ratings on coarse criteria (Cohen's \kappa=0.70) but break down on fine-grained, part-level judgment (\kappa=0.10), validating the paradigm in its strong regime while clarifying its limits.
Figures
Figures from the paper (5 more)
Reference graph
Works this paper leans on
-
[1]
Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
IMoRe: Implicit Program-Guided Reasoning for Human Motion Q&A , author=. Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
-
[2]
arXiv preprint arXiv:2602.17768 , year=
KPM-Bench: A Kinematic Parsing Motion Benchmark for Fine-grained Motion-centric Video Understanding , author=. arXiv preprint arXiv:2602.17768 , year=
-
[3]
Proceedings of the 33rd ACM International Conference on Multimedia , pages=
Towards Fine-Grained Human Motion Video Captioning , author=. Proceedings of the 33rd ACM International Conference on Multimedia , pages=
-
[4]
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding , author=. arXiv preprint arXiv:2605.19846 , year=
-
[5]
Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
Motionbench: Benchmarking and improving fine-grained video motion understanding for vision language models , author=. Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
-
[6]
IEEE Transactions on Pattern Analysis and Machine Intelligence , year=
Motionllm: Understanding human behaviors from human motions and videos , author=. IEEE Transactions on Pattern Analysis and Machine Intelligence , year=
-
[7]
Li, Chuqiao and Xie, Xianghui and Cao, Yong and Geiger, Andreas and Pons-Moll, Gerard , booktitle=
-
[8]
and Chandrasekaran, Arjun and Athanasiou, Nikos and Quiros-Ramirez, Alejandra and Black, Michael J
Punnakkal, Abhinanda R. and Chandrasekaran, Arjun and Athanasiou, Nikos and Quiros-Ramirez, Alejandra and Black, Michael J. , booktitle =. 2021 , doi =
2021
Show all 300 references
-
[9]
Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , month =
Xiao, Lixing and Lu, Shunlin and Pi, Huaijin and Fan, Ke and Pan, Liang and Zhou, Yueer and Feng, Ziyong and Zhou, Xiaowei and Peng, Sida and Wang, Jingbo , title =. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , month =. 2025 , pages =
2025
-
[10]
Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
Rethinking diffusion for text-driven human motion generation: Redundant representations, evaluation, and masked autoregression , author=. Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
-
[11]
arXiv preprint arXiv:2603.13500 , year=
ActionPlan: Future-Aware Streaming Motion Synthesis via Frame-Level Action Planning , author=. arXiv preprint arXiv:2603.13500 , year=
-
[12]
European Conference on Computer Vision , pages=
Como: Controllable motion generation through language guided pose code editing , author=. European Conference on Computer Vision , pages=. 2024 , organization=
2024
-
[13]
The Fourteenth International Conference on Learning Representations , year=
The Quest for Generalizable Motion Generation: Data, Model, and Evaluation , author=. The Fourteenth International Conference on Learning Representations , year=
-
[14]
and Pons-Moll, Gerard and Black, Michael J
Mahmood, Naureen and Ghorbani, Nima and Troje, Nikolaus F. and Pons-Moll, Gerard and Black, Michael J. , booktitle =. 2019 , month_numeric =
2019
-
[15]
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =
Guo, Chuan and Zou, Shihao and Zuo, Xinxin and Wang, Sen and Ji, Wei and Li, Xingyu and Cheng, Li , title =. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =. 2022 , pages =
2022
-
[16]
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Momask: Generative masked modeling of 3d human motions , author=. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
-
[17]
arXiv preprint arXiv:2209.14916 , year=
Human motion diffusion model , author=. arXiv preprint arXiv:2209.14916 , year=
-
[18]
Advances in Neural Information Processing Systems , volume=
Motiongpt: Human motion as a foreign language , author=. Advances in Neural Information Processing Systems , volume=. 2023 , eprint=
2023
-
[19]
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=
T2M-GPT: Generating Human Motion from Textual Descriptions with Discrete Representations , author=. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=
-
[20]
International Conference on Machine Learning , pages=
Motion question answering via modular motion programs , author=. International Conference on Machine Learning , pages=. 2023 , organization=. 2305.08953 , archivePrefix=
2023
-
[21]
International Conference on Machine Learning , pages=
Motion question answering via modular motion programs , author=. International Conference on Machine Learning , pages=. 2023 , organization=
2023
-
[22]
The Eleventh International Conference on Learning Representations , year=
Human Motion Diffusion Model , author=. The Eleventh International Conference on Learning Representations , year=
-
[23]
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Executing Your Commands via Motion Diffusion in Latent Space , author=. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=. 2023 , eprint=
2023
-
[24]
arXiv preprint arXiv:2405.20340 , year=
MotionLLM: Understanding Human Behaviors from Human Motions and Videos , author=. arXiv preprint arXiv:2405.20340 , year=. 2405.20340 , archivePrefix=
-
[25]
arXiv preprint arXiv:2405.17013 , year=
Motion-Agent: A Conversational Framework for Human Motion Generation with LLMs , author=. arXiv preprint arXiv:2405.17013 , year=. 2405.17013 , archivePrefix=
-
[26]
arXiv preprint arXiv:2411.17335 , year=
VersatileMotion: A Unified Framework for Motion Synthesis and Comprehension , author=. arXiv preprint arXiv:2411.17335 , year=. 2411.17335 , archivePrefix=
-
[27]
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , year=
MG-MotionLLM: A Unified Framework for Motion Comprehension and Generation across Multiple Granularities , author=. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , year=. 2504.02478 , archivePrefix=
-
[28]
Advances in Neural Information Processing Systems , year=
Visual Instruction Tuning , author=. Advances in Neural Information Processing Systems , year=. 2304.08485 , archivePrefix=
-
[29]
Proceedings of the Annual Meeting of the Association for Computational Linguistics , year=
Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models , author=. Proceedings of the Annual Meeting of the Association for Computational Linguistics , year=. 2306.05424 , archivePrefix=
-
[30]
arXiv preprint arXiv:2511.21631 , year=
Qwen3-vl technical report , author=. arXiv preprint arXiv:2511.21631 , year=
-
[31]
arXiv preprint arXiv:2503.14935 , year=
FAVOR-Bench: A Comprehensive Benchmark for Fine-Grained Video Motion Understanding , author=. arXiv preprint arXiv:2503.14935 , year=
-
[32]
Proceedings of the Conference on Empirical Methods in Natural Language Processing , year=
G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment , author=. Proceedings of the Conference on Empirical Methods in Natural Language Processing , year=. 2303.16634 , archivePrefix=
-
[33]
arXiv preprint arXiv:2401.06591 , year=
Prometheus-Vision: Vision-Language Model as a Judge for Fine-Grained Evaluation , author=. arXiv preprint arXiv:2401.06591 , year=. 2401.06591 , archivePrefix=
-
[34]
International Conference on Machine Learning , year=
MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark , author=. International Conference on Machine Learning , year=. 2402.04788 , archivePrefix=
-
[35]
arXiv preprint arXiv:2307.15818 , year=
RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control , author=. arXiv preprint arXiv:2307.15818 , year=. 2307.15818 , archivePrefix=
-
[36]
arXiv preprint arXiv:2406.09246 , year=
OpenVLA: An Open-Source Vision-Language-Action Model , author=. arXiv preprint arXiv:2406.09246 , year=. 2406.09246 , archivePrefix=
-
[37]
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Generating Diverse and Natural 3D Human Motions from Text , author=. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
-
[38]
Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
Remodiffuse: Retrieval-augmented motion diffusion model , author=. Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
-
[39]
2023 , journal =
FineMoGen: Fine-Grained Spatio-Temporal Motion Generation and Editing , author =. 2023 , journal =
2023
-
[40]
European Conference on Computer Vision , year=
MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model , author=. European Conference on Computer Vision , year=. 2404.19759 , archivePrefix=
-
[41]
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Avatargpt: All-in-one framework for motion understanding planning generation and beyond , author=. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
-
[42]
arXiv preprint arXiv:2408.03326 , year=
LLaVA-OneVision: Easy Visual Task Transfer , author=. arXiv preprint arXiv:2408.03326 , year=. 2408.03326 , archivePrefix=
-
[43]
arXiv preprint arXiv:2501.13106 , year=
Videollama 3: Frontier multimodal foundation models for image and video understanding , author=. arXiv preprint arXiv:2501.13106 , year=
-
[44]
Advances in Neural Information Processing Systems , year=
EgoSchema: A Diagnostic Benchmark for Very Long-form Video Language Understanding , author=. Advances in Neural Information Processing Systems , year=. 2308.09126 , archivePrefix=
-
[45]
Findings of the Association for Computational Linguistics: ACL 2024 , pages=
TempCompass: Do Video LLMs Really Understand Videos? , author=. Findings of the Association for Computational Linguistics: ACL 2024 , pages=. 2024 , eprint=
2024
-
[46]
arXiv preprint arXiv:2405.21075 , year=
Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis , author=. arXiv preprint arXiv:2405.21075 , year=. 2405.21075 , archivePrefix=
-
[47]
Advances in neural information processing systems , volume=
Judging llm-as-a-judge with mt-bench and chatbot arena , author=. Advances in neural information processing systems , volume=
-
[48]
Robotics: Science and Systems , year=
Octo: An Open-Source Generalist Robot Policy , author=. Robotics: Science and Systems , year=. 2405.12213 , archivePrefix=
-
[49]
arXiv preprint arXiv:2410.24164 , year=
_0 : A Vision-Language-Action Flow Model for General Robot Control , author=. arXiv preprint arXiv:2410.24164 , year=. 2410.24164 , archivePrefix=
-
[50]
arXiv preprint arXiv:2410.07864 , year=
RDT-1B: A Diffusion Foundation Model for Bimanual Manipulation , author=. arXiv preprint arXiv:2410.07864 , year=. 2410.07864 , archivePrefix=
-
[51]
Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , month =
Xie, Xianghui and Lessen, Jan Eric and Pons-Moll, Gerard , title =. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , month =. 2025 , pages =
2025
-
[52]
Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024
2024
-
[53]
Investigating radicalisation indicators in online extremist communities
De Kock, Christine and Hovy, Eduard. Investigating radicalisation indicators in online extremist communities. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.1
2024 doi
-
[54]
Detection of Conspiracy Theories Beyond Keyword Bias in G erman-Language Telegram Using Large Language Models
Pustet, Milena and Steffen, Elisabeth and Mihaljevic, Helena. Detection of Conspiracy Theories Beyond Keyword Bias in G erman-Language Telegram Using Large Language Models. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.2
2024 doi
-
[55]
E ko H ate: Abusive Language and Hate Speech Detection for Code-switched Political Discussions on N igerian T witter
Ilevbare, Comfort and Alabi, Jesujoba and Adelani, David Ifeoluwa and Bakare, Firdous and Abiola, Oluwatoyin and Adeyemo, Oluwaseyi. E ko H ate: Abusive Language and Hate Speech Detection for Code-switched Political Discussions on N igerian T witter. Proceedings of the 8th Wor...
2024 doi
-
[56]
A Study of the Class Imbalance Problem in Abusive Language Detection
Zhang, Yaqi and Hangya, Viktor and Fraser, Alexander. A Study of the Class Imbalance Problem in Abusive Language Detection. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.4
2024 doi
-
[57]
H ausa H ate: An Expert Annotated Corpus for H ausa Hate Speech Detection
Vargas, Francielle and Guimar \ a es, Samuel and Muhammad, Shamsuddeen Hassan and Alves, Diego and Ahmad, Ibrahim Said and Abdulmumin, Idris and Mohamed, Diallo and Pardo, Thiago and Benevenuto, Fabr \' cio. H ausa H ate: An Expert Annotated Corpus for H ausa Hate Speech Detec...
2024 doi
-
[58]
VIDA : The Visual Incel Data Archive
Anastasi, Selenia and Schneider, Florian and Biemann, Chris and Fischer, Tim. VIDA : The Visual Incel Data Archive. A Theory-oriented Annotated Dataset To Enhance Hate Detection Through Visual Culture. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024...
2024 doi
-
[59]
Towards a Unified Framework for Adaptable Problematic Content Detection via Continual Learning
Omrani, Ali and Salkhordeh Ziabari, Alireza and Golazizian, Preni and Sorensen, Jeffrey and Dehghani, Morteza. Towards a Unified Framework for Adaptable Problematic Content Detection via Continual Learning. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024)....
2024 doi
-
[60]
From Linguistics to Practice: a Case Study of Offensive Language Taxonomy in H ebrew
Liebeskind, Chaya and Litvak, Marina and Vanetik, Natalia. From Linguistics to Practice: a Case Study of Offensive Language Taxonomy in H ebrew. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.8
2024 doi
-
[61]
Estimating the Emotion of Disgust in G reek Parliament Records
Lislevand, Vanessa and Pavlopoulos, John and Louridas, Panos and Dritsa, Konstantina. Estimating the Emotion of Disgust in G reek Parliament Records. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.9
2024 doi
-
[62]
Simple LLM based Approach to Counter Algospeak
Fillies, Jan and Paschke, Adrian. Simple LLM based Approach to Counter Algospeak. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.10
2024 doi
-
[63]
Harnessing Personalization Methods to Identify and Predict Unreliable Information Spreader Behavior
Ashraf, Shaina and Gruschka, Fabio and Flek, Lucie and Welch, Charles. Harnessing Personalization Methods to Identify and Predict Unreliable Information Spreader Behavior. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.11
2024 doi
-
[64]
Robust Safety Classifier Against Jailbreaking Attacks: Adversarial Prompt Shield
Kim, Jinhwa and Derakhshan, Ali and Harris, Ian. Robust Safety Classifier Against Jailbreaking Attacks: Adversarial Prompt Shield. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.12
2024 doi
-
[65]
Improving aggressiveness detection using a data augmentation technique based on a Diffusion Language Model
Reyes-Ram \' rez, Antonio and Arag \'o n, Mario and S \'a nchez-Vega, Fernando and L \'o pez-Monroy, Adrian. Improving aggressiveness detection using a data augmentation technique based on a Diffusion Language Model. Proceedings of the 8th Workshop on Online Abuse and Harms (W...
2024 doi
-
[66]
The M exican Gayze: A Computational Analysis of the Attitudes towards the LGBT + Population in M exico on Social Media Across a Decade
Andersen, Scott and Ojeda-Trueba, Segio-Luis and V \'a squez, Juan and Bel-Enguix, Gemma. The M exican Gayze: A Computational Analysis of the Attitudes towards the LGBT + Population in M exico on Social Media Across a Decade. Proceedings of the 8th Workshop on Online Abuse and...
2024 doi
-
[67]
X -posing Free Speech: Examining the Impact of Moderation Relaxation on Online Social Networks
Arun, Arvindh and Chhatani, Saurav and An, Jisun and Kumaraguru, Ponnurangam. X -posing Free Speech: Examining the Impact of Moderation Relaxation on Online Social Networks. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.15
2024 doi
-
[68]
The Uli Dataset: An Exercise in Experience Led Annotation of o GBV
Arora, Arnav and Jinadoss, Maha and Arora, Cheshta and George, Denny and Brindaalakshmi and Khan, Haseena and Rawat, Kirti and Div and Ritash and Mathur, Seema. The Uli Dataset: An Exercise in Experience Led Annotation of o GBV. Proceedings of the 8th Workshop on Online Abuse ...
2024 doi
-
[69]
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
Nirmal, Ayushi and Bhattacharjee, Amrita and Sheth, Paras and Liu, Huan. Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.17
2024 doi
-
[70]
A B ayesian Quantification of Aporophobia and the Aggravating Effect of Low -- Wealth Contexts on Stigmatization
Brate, Ryan and Van Erp, Marieke and Van Den Bosch, Antal. A B ayesian Quantification of Aporophobia and the Aggravating Effect of Low -- Wealth Contexts on Stigmatization. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.18
2024 doi
-
[71]
Toxicity Classification in U krainian
Dementieva, Daryna and Khylenko, Valeriia and Babakov, Nikolay and Groh, Georg. Toxicity Classification in U krainian. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.19
2024 doi
-
[72]
A Strategy Labelled Dataset of Counterspeech
Poudhar, Aashima and Konstas, Ioannis and Abercrombie, Gavin. A Strategy Labelled Dataset of Counterspeech. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.20
2024 doi
-
[73]
Improving Covert Toxicity Detection by Retrieving and Generating References
Lee, Dong-Ho and Cho, Hyundong and Jin, Woojeong and Moon, Jihyung and Park, Sungjoon and R. Improving Covert Toxicity Detection by Retrieving and Generating References. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.21
2024 doi
-
[74]
Subjective Isms? On the Danger of Conflating Hate and Offence in Abusive Language Detection
Cercas Curry, Amanda and Abercrombie, Gavin and Talat, Zeerak. Subjective Isms? On the Danger of Conflating Hate and Offence in Abusive Language Detection. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v1/2024.woah-1.22
2024 doi
-
[75]
From Languages to Geographies: Towards Evaluating Cultural Bias in Hate Speech Datasets
Tonneau, Manuel and Liu, Diyi and Fraiberger, Samuel and Schroeder, Ralph and Hale, Scott and R. From Languages to Geographies: Towards Evaluating Cultural Bias in Hate Speech Datasets. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. doi:10.18653/v...
2024 doi
-
[76]
SGH ate C heck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of S ingapore
Ng, Ri Chi and Prakash, Nirmalendu and Hee, Ming Shan and Choo, Kenny Tsu Wei and Lee, Roy Ka-wei. SGH ate C heck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of S ingapore. Proceedings of the 8th Workshop on Online Abuse and Harms (WOAH 2024). 2024. d...
2024 doi
-
[77]
Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[78]
Correcting Challenging F innish Learner Texts With Claude, GPT -3.5 and GPT -4 Large Language Models
Creutz, Mathias. Correcting Challenging F innish Learner Texts With Claude, GPT -3.5 and GPT -4 Large Language Models. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[79]
Context-aware Adversarial Attack on Named Entity Recognition
Chen, Shuguang and Neves, Leonardo and Solorio, Thamar. Context-aware Adversarial Attack on Named Entity Recognition. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[80]
Effects of different types of noise in user-generated reviews on human and machine translations including C hat GPT
Popovic, Maja and Lapshinova-Koltunski, Ekaterina and Koponen, Maarit. Effects of different types of noise in user-generated reviews on human and machine translations including C hat GPT. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[81]
Stanceosaurus 2.0 - Classifying Stance Towards R ussian and S panish Misinformation
Lavrouk, Anton and Ligon, Ian and Zheng, Jonathan and Naous, Tarek and Xu, Wei and Ritter, Alan. Stanceosaurus 2.0 - Classifying Stance Towards R ussian and S panish Misinformation. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[82]
and Shahariar, G
Elahi, Kazi and Rahman, Tasnuva and Shahriar, Shakil and Sarker, Samir and Shawon, Md. and Shahariar, G. M. A Comparative Analysis of Noise Reduction Methods in Sentiment Analysis on Noisy B angla Texts. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT...
2024
-
[83]
Label Supervised Contrastive Learning for Imbalanced Text Classification in E uclidean and Hyperbolic Embedding Spaces
Khalid, Baber and Dai, Shuyang and Taghavi, Tara and Lee, Sungjin. Label Supervised Contrastive Learning for Imbalanced Text Classification in E uclidean and Hyperbolic Embedding Spaces. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[84]
M aint N orm: A corpus and benchmark model for lexical normalisation and masking of industrial maintenance short text
Bikaun, Tyler and Hodkiewicz, Melinda and Liu, Wei. M aint N orm: A corpus and benchmark model for lexical normalisation and masking of industrial maintenance short text. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[85]
The Effects of Data Quality on Named Entity Recognition
Bhadauria, Divya and Sierra M \'u nera, Alejandro and Krestel, Ralf. The Effects of Data Quality on Named Entity Recognition. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[86]
Topic Bias in Emotion Classification
Wegge, Maximilian and Klinger, Roman. Topic Bias in Emotion Classification. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[87]
Stars Are All You Need: A Distantly Supervised Pyramid Network for Unified Sentiment Analysis
Li, Wenchang and Chen, Yixing and Zheng, Shuang and Wang, Lei and Lalor, John. Stars Are All You Need: A Distantly Supervised Pyramid Network for Unified Sentiment Analysis. Proceedings of the Ninth Workshop on Noisy and User-generated Text (W-NUT 2024). 2024
2024
-
[88]
Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[89]
Towards Disfluency Annotated Corpora for I ndian Languages
Kochar, Chayan and Mujadia, Vandan Vasantlal and Mishra, Pruthwik and Sharma, Dipti Misra. Towards Disfluency Annotated Corpora for I ndian Languages. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[90]
E mo M ix-3 L : A Code-Mixed Dataset for B angla- E nglish- H indi for Emotion Detection
Raihan, Nishat and Goswami, Dhiman and Mahmud, Antara and Anastasopoulos, Antonios and Zampieri, Marcos. E mo M ix-3 L : A Code-Mixed Dataset for B angla- E nglish- H indi for Emotion Detection. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[91]
and Buitelaar, Paul and McCrae, John P
Rani, Priya and Negi, Gaurav and Jha, Saroj and Suryawanshi, Shardul and Ojha, Atul Kr. and Buitelaar, Paul and McCrae, John P. Findings of the WILDRE Shared Task on Code-mixed Less-resourced Sentiment Analysis for I ndo- A ryan Languages. Proceedings of the 7th Workshop on In...
2024
-
[92]
Multilingual Bias Detection and Mitigation for I ndian Languages
Maity, Ankita and Sharma, Anubhav and Dhar, Rudra and Abhishek, Tushar and Gupta, Manish and Varma, Vasudeva. Multilingual Bias Detection and Mitigation for I ndian Languages. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[93]
Dharma \'s \=a stra Informatics: Concept Mining System for Socio-Cultural Facet in A ncient I ndia
Nigam, Arooshi and Chandra, Subhash. Dharma \'s \=a stra Informatics: Concept Mining System for Socio-Cultural Facet in A ncient I ndia. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[94]
Exploring News Summarization and Enrichment in a Highly Resource-Scarce I ndian Language: A Case Study of Mizo
Bala, Abhinaba and Urlana, Ashok and Mishra, Rahul and Krishnamurthy, Parameswari. Exploring News Summarization and Enrichment in a Highly Resource-Scarce I ndian Language: A Case Study of Mizo. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[95]
Finding the Causality of an Event in News Articles
Lalitha Devi, Sobha and RK Rao, Pattabhi. Finding the Causality of an Event in News Articles. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[96]
Creating Corpus of Low Resource I ndian Languages for Natural Language Processing: Challenges and Opportunities
Dongare, Pratibha. Creating Corpus of Low Resource I ndian Languages for Natural Language Processing: Challenges and Opportunities. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[97]
FZZG at WILDRE -7: Fine-tuning Pre-trained Models for Code-mixed, Less-resourced Sentiment Analysis
Thakkar, Gaurish and Tadi \'c , Marko and Mikelic Preradovic, Nives. FZZG at WILDRE -7: Fine-tuning Pre-trained Models for Code-mixed, Less-resourced Sentiment Analysis. Proceedings of the 7th Workshop on Indian Language Data: Resources and Evaluation. 2024
2024
-
[98]
MLI nitiative@ WILDRE 7: Hybrid Approaches with Large Language Models for Enhanced Sentiment Analysis in Code-Switched and Code-Mixed Texts
Veeramani, Hariram and Thapa, Surendrabikram and Naseem, Usman. MLI nitiative@ WILDRE 7: Hybrid Approaches with Large Language Models for Enhanced Sentiment Analysis in Code-Switched and Code-Mixed Texts. Proceedings of the 7th Workshop on Indian Language Data: Resources and E...
2024
-
[99]
Aalamaram: A Large-Scale Linguistically Annotated Treebank for the T amil Language
Abirami, A M and Leong, Wei Qi and Rengarajan, Hamsawardhini and Anitha, D and Suganya, R and Singh, Himanshu and Sarveswaran, Kengatharaiyer and Tjhi, William Chandra and Shah, Rajiv Ratn. Aalamaram: A Large-Scale Linguistically Annotated Treebank for the T amil Language. Pro...
2024
-
[100]
Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[101]
Enhanced Financial Sentiment Analysis and Trading Strategy Development Using Large Language Models
Kirtac, Kemal and Germano, Guido. Enhanced Financial Sentiment Analysis and Trading Strategy Development Using Large Language Models. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[102]
SEC : Context-Aware Metric Learning for Efficient Emotion Recognition in Conversation
Gendron, Barbara and Guibon, Ga. SEC : Context-Aware Metric Learning for Efficient Emotion Recognition in Conversation. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[103]
Modeling Complex Interactions in Long Documents for Aspect-Based Sentiment Analysis
Yan, Zehong and Hsu, Wynne and Lee, Mong-Li and Bartram-Shaw, David. Modeling Complex Interactions in Long Documents for Aspect-Based Sentiment Analysis. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[104]
Hierarchical Adversarial Correction to Mitigate Identity Term Bias in Toxicity Detection
Sch. Hierarchical Adversarial Correction to Mitigate Identity Term Bias in Toxicity Detection. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[105]
A Systematic Analysis on the Temporal Generalization of Language Models in Social Media
Ushio, Asahi and Camacho-Collados, Jose. A Systematic Analysis on the Temporal Generalization of Language Models in Social Media. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[106]
LL a MA -Based Models for Aspect-Based Sentiment Analysis
S m \' d, Jakub and Priban, Pavel and Kral, Pavel. LL a MA -Based Models for Aspect-Based Sentiment Analysis. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[107]
A Multi-Faceted NLP Analysis of Misinformation Spreaders in T witter
Antypas, Dimosthenis and Preece, Alun and Camacho-Collados, Jose. A Multi-Faceted NLP Analysis of Misinformation Spreaders in T witter. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[108]
Entity-Level Sentiment: More than the Sum of Its Parts
R nningstad, Egil and Klinger, Roman and vrelid, Lilja and Velldal, Erik. Entity-Level Sentiment: More than the Sum of Its Parts. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[109]
MBIAS : Mitigating Bias in Large Language Models While Retaining Context
Raza, Shaina and Raval, Ananya and Chatrath, Veronica. MBIAS : Mitigating Bias in Large Language Models While Retaining Context. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[110]
Polarization of Autonomous Generative AI Agents Under Echo Chambers
Ohagi, Masaya. Polarization of Autonomous Generative AI Agents Under Echo Chambers. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[111]
Know Thine Enemy: Adaptive Attacks on Misinformation Detection Using Reinforcement Learning
Przyby a, Piotr and McGill, Euan and Saggion, Horacio. Know Thine Enemy: Adaptive Attacks on Misinformation Detection Using Reinforcement Learning. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[112]
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models
Zhu, Xiliang and Gardiner, Shayna and Rold \'a n, Tere and Rossouw, David. The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \&...
2024
-
[113]
Guiding Sentiment Analysis with Hierarchical Text Clustering: Analyzing the G erman X / T witter Discourse on Face Masks in the 2020 COVID -19 Pandemic
Wehrli, Silvan and Ezekannagha, Chisom and Hattab, Georges and Boender, Tamara and Arnrich, Bert and Irrgang, Christopher. Guiding Sentiment Analysis with Hierarchical Text Clustering: Analyzing the G erman X / T witter Discourse on Face Masks in the 2020 COVID -19 Pandemic. P...
2020
-
[114]
Emotion Identification for F rench in Written Texts: Considering Modes of Emotion Expression as a Step Towards Text Complexity Analysis
\'E tienne, Aline and Battistelli, Delphine and Lecorv \'e , Gw \'e nol \'e. Emotion Identification for F rench in Written Texts: Considering Modes of Emotion Expression as a Step Towards Text Complexity Analysis. Proceedings of the 14th Workshop on Computational Approaches to...
2024
-
[115]
Comparing Tools for Sentiment Analysis of D anish Literature from Hymns to Fairy Tales: Low-Resource Language and Domain Challenges
Feldkamp, Pascale and Kostkan, Jan and Overgaard, Ea and Jacobsen, Mia and Bizzoni, Yuri. Comparing Tools for Sentiment Analysis of D anish Literature from Hymns to Fairy Tales: Low-Resource Language and Domain Challenges. Proceedings of the 14th Workshop on Computational Appr...
2024
-
[116]
Multi-Target User Stance Discovery on R eddit
Steel, Benjamin and Ruths, Derek. Multi-Target User Stance Discovery on R eddit. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[117]
Subjectivity Detection in E nglish News using Large Language Models
Shokri, Mohammad and Sharma, Vivek and Filatova, Elena and Jain, Shweta and Levitan, Sarah. Subjectivity Detection in E nglish News using Large Language Models. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[118]
Monitoring Depression Severity and Symptoms in User-Generated Content: An Annotation Scheme and Guidelines
Alhamed, Falwah and Bendayan, Rebecca and Ive, Julia and Specia, Lucia. Monitoring Depression Severity and Symptoms in User-Generated Content: An Annotation Scheme and Guidelines. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Socia...
2024
-
[119]
R ide KE : Leveraging Low-resource T witter User-generated Content for Sentiment and Emotion Detection on Code-switched RHS Dataset
Etori, Naome and Gini, Maria. R ide KE : Leveraging Low-resource T witter User-generated Content for Sentiment and Emotion Detection on Code-switched RHS Dataset. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[120]
POL ygraph: P olish Fake News Dataset
Dzienisiewicz, Daniel and Grali \'n ski, Filip and Jab o \'n ski, Piotr and Kubis, Marek and Sk \'o rzewski, Pawe and Wierzchon, Piotr. POL ygraph: P olish Fake News Dataset. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Med...
2024
-
[121]
Exploring Language Models to Analyze Market Demand Sentiments from News
Dasgupta, Tirthankar and Sinha, Manjira. Exploring Language Models to Analyze Market Demand Sentiments from News. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[122]
Impact of Decoding Methods on Human Alignment of Conversational LLM s
Furniturewala, Shaz and Jaidka, Kokil and Sharma, Yashvardhan. Impact of Decoding Methods on Human Alignment of Conversational LLM s. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[123]
Loneliness Episodes: A J apanese Dataset for Loneliness Detection and Analysis
Fujikawa, Naoya and Toan, Nguyen and Ito, Kazuhiro and Wakamiya, Shoko and Aramaki, Eiji. Loneliness Episodes: A J apanese Dataset for Loneliness Detection and Analysis. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media An...
2024
-
[124]
Estimation of Happiness Changes through Longitudinal Analysis of Employees ' Texts
Hayashi, Junko and Ito, Kazuhiro and Manabe, Masae and Watanabe, Yasushi and Nakayama, Masataka and Uchida, Yukiko and Wakamiya, Shoko and Aramaki, Eiji. Estimation of Happiness Changes through Longitudinal Analysis of Employees ' Texts. Proceedings of the 14th Workshop on Com...
2024
-
[125]
Subjectivity Theory vs
Savinova, Elena and Hoek, Jet. Subjectivity Theory vs. Speaker Intuitions: Explaining the Results of a Subjectivity Regressor Trained on Native Speaker Judgements. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[126]
Andrew and Hovy, Dirk
Soni, Nikita and Balasubramanian, Niranjan and Schwartz, H. Andrew and Hovy, Dirk. Comparing Pre-trained Human Language Models: Is it Better with Human Context as Groups, Individual Traits, or Both?. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity,...
2024
-
[127]
LLM s for Targeted Sentiment in News Headlines: Exploring the Descriptive-Prescriptive Dilemma
Juro s , Jana and Majer, Laura and Snajder, Jan. LLM s for Targeted Sentiment in News Headlines: Exploring the Descriptive-Prescriptive Dilemma. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[128]
Context is Important in Depressive Language: A Study of the Interaction Between the Sentiments and Linguistic Markers in R eddit Discussions
Sharma, Neha and Sirts, Kairit. Context is Important in Depressive Language: A Study of the Interaction Between the Sentiments and Linguistic Markers in R eddit Discussions. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Medi...
2024
-
[129]
To Aggregate or Not to Aggregate
Kurniawan, Kemal and Mistica, Meladel and Baldwin, Timothy and Lau, Jey Han. To Aggregate or Not to Aggregate. That is the Question: A Case Study on Annotation Subjectivity in Span Prediction. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentim...
2024
-
[130]
Findings of WASSA 2024 Shared Task on Empathy and Personality Detection in Interactions
Giorgi, Salvatore and Sedoc, Jo \ a o and Barriere, Valentin and Tafreshi, Shabnam. Findings of WASSA 2024 Shared Task on Empathy and Personality Detection in Interactions. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media...
2024
-
[131]
RU at WASSA 2024 Shared Task: Task-Aligned Prompt for Predicting Empathy and Distress
Kong, Haein and Moon, Seonghyeon. RU at WASSA 2024 Shared Task: Task-Aligned Prompt for Predicting Empathy and Distress. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[132]
Chinchunmei at WASSA 2024 Empathy and Personality Shared Task: Boosting LLM ' s Prediction with Role-play Augmentation and Contrastive Reasoning Calibration
Li, Tian and Rusnachenko, Nicolay and Liang, Huizhi. Chinchunmei at WASSA 2024 Empathy and Personality Shared Task: Boosting LLM ' s Prediction with Role-play Augmentation and Contrastive Reasoning Calibration. Proceedings of the 14th Workshop on Computational Approaches to Su...
2024
-
[133]
Empathify at WASSA 2024 Empathy and Personality Shared Task: Contextualizing Empathy with a BERT -Based Context-Aware Approach for Empathy Detection
Numano. Empathify at WASSA 2024 Empathy and Personality Shared Task: Contextualizing Empathy with a BERT -Based Context-Aware Approach for Empathy Detection. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[134]
Huang, Liting and Liang, Huizhi. Zhenmei at WASSA -2024 Empathy and Personality Shared Track 2 Incorporating P earson Correlation Coefficient as a Regularization Term for Enhanced Empathy and Emotion Prediction in Conversational Turns. Proceedings of the 14th Workshop on Compu...
2024
-
[135]
Empaths at WASSA 2024 Empathy and Personality Shared Task: Turn-Level Empathy Prediction Using Psychological Indicators
Furniturewala, Shaz and Jaidka, Kokil. Empaths at WASSA 2024 Empathy and Personality Shared Task: Turn-Level Empathy Prediction Using Psychological Indicators. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[136]
NU at WASSA 2024 Empathy and Personality Shared Task: Enhancing Personality Predictions with Knowledge Graphs; A Graphical Neural Network and L ight GBM Ensemble Approach
Osei-Brefo, Emmanuel and Liang, Huizhi. NU at WASSA 2024 Empathy and Personality Shared Task: Enhancing Personality Predictions with Knowledge Graphs; A Graphical Neural Network and L ight GBM Ensemble Approach. Proceedings of the 14th Workshop on Computational Approaches to S...
2024
-
[137]
Daisy at WASSA 2024 Empathy and Personality Shared Task: A Quick Exploration on Emotional Pattern of Empathy and Distress
Chevi, Rendi and Aji, Alham. Daisy at WASSA 2024 Empathy and Personality Shared Task: A Quick Exploration on Emotional Pattern of Empathy and Distress. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[138]
WASSA 2024 Shared Task: Enhancing Emotional Intelligence with Prompts
Churina, Svetlana and Verma, Preetika and Tripathy, Suchismita. WASSA 2024 Shared Task: Enhancing Emotional Intelligence with Prompts. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[139]
Yang, Huiyu and Huang, Liting and Li, Tian and Rusnachenko, Nicolay and Liang, Huizhi. hyy33 at WASSA 2024 Empathy and Personality Shared Task: Using the C ombined L oss and FGM for Enhancing BERT -based Models in Emotion and Empathy Prediction from Conversation Turns. Proceed...
2024
-
[140]
Frick, Raphael and Steinebach, Martin. Fraunhofer SIT at WASSA 2024 Empathy and Personality Shared Task: Use of Sentiment Transformers and Data Augmentation With Fuzzy Labels to Predict Emotional Reactions in Conversations and Essays. Proceedings of the 14th Workshop on Comput...
2024
-
[141]
and Parde, Natalie
Lee, Gyeongeun and Wang, Zhu and Ravi, Sathya N. and Parde, Natalie. E mpathetic FIG at WASSA 2024 Empathy and Personality Shared Task: Predicting Empathy and Emotion in Conversations with Figurative Language. Proceedings of the 14th Workshop on Computational Approaches to Sub...
2024
-
[142]
C on T ext at WASSA 2024 Empathy and Personality Shared Task: History-Dependent Embedding Utterance Representations for Empathy and Emotion Prediction in Conversations
Pereira, Patr \' cia and Moniz, Helena and Carvalho, Joao Paulo. C on T ext at WASSA 2024 Empathy and Personality Shared Task: History-Dependent Embedding Utterance Representations for Empathy and Emotion Prediction in Conversations. Proceedings of the 14th Workshop on Computa...
2024
-
[143]
Findings of the WASSA 2024 EXALT shared task on Explainability for Cross-Lingual Emotion in Tweets
Maladry, Aaron and Singh, Pranaydeep and Lefever, Els. Findings of the WASSA 2024 EXALT shared task on Explainability for Cross-Lingual Emotion in Tweets. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[144]
Cross-lingual Emotion Detection through Large Language Models
Kadiyala, Ram Mohan Rao. Cross-lingual Emotion Detection through Large Language Models. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[145]
Knowledge Distillation from Monolingual to Multilingual Models for Intelligent and Interpretable Multilingual Emotion Detection
Wang, Yuqi and Wang, Zimu and Han, Nijia and Wang, Wei and Chen, Qi and Zhang, Haiyang and Pan, Yushan and Nguyen, Anh. Knowledge Distillation from Monolingual to Multilingual Models for Intelligent and Interpretable Multilingual Emotion Detection. Proceedings of the 14th Work...
2024
-
[146]
HITSZ - HLT at WASSA -2024 Shared Task 2: Language-agnostic Multi-task Learning for Explainability of Cross-lingual Emotion Detection
Xiong, Feng and Wang, Jun and Tu, Geng and Xu, Ruifeng. HITSZ - HLT at WASSA -2024 Shared Task 2: Language-agnostic Multi-task Learning for Explainability of Cross-lingual Emotion Detection. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentimen...
2024
-
[147]
UWB at WASSA -2024 Shared Task 2: Cross-lingual Emotion Detection
S m \' d, Jakub and P r ib \'a n , Pavel and Kr \'a l, Pavel. UWB at WASSA -2024 Shared Task 2: Cross-lingual Emotion Detection. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media Analysis. 2024
2024
-
[148]
PCICUNAM at WASSA 2024: Cross-lingual Emotion Detection Task with Hierarchical Classification and Weighted Loss Functions
V \'a zquez-Osorio, Jes \'u s and Sierra, Gerardo and G \'o mez-Adorno, Helena and Bel-Enguix, Gemma. PCICUNAM at WASSA 2024: Cross-lingual Emotion Detection Task with Hierarchical Classification and Weighted Loss Functions. Proceedings of the 14th Workshop on Computational Ap...
2024
-
[149]
TEII : Think, Explain, Interact and Iterate with Large Language Models to Solve Cross-lingual Emotion Detection
Cheng, Long and Shao, Qihao and Zhao, Christine and Bi, Sheng and Levow, Gina-Anne. TEII : Think, Explain, Interact and Iterate with Large Language Models to Solve Cross-lingual Emotion Detection. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Se...
2024
-
[150]
NYCU - NLP at EXALT 2024: Assembling Large Language Models for Cross-Lingual Emotion and Trigger Detection
Lin, Tzu-Mi and Xu, Zhe-Yu and Zhou, Jian-Yu and Lee, Lung-Hao. NYCU - NLP at EXALT 2024: Assembling Large Language Models for Cross-Lingual Emotion and Trigger Detection. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media ...
2024
-
[151]
Effectiveness of Scalable Monolingual Data and Trigger Words Prompting on Cross-Lingual Emotion Detection Task
Cheng, Yao-Fei and Hong, Jeongyeob and Wang, Andrew and Silva, Anita and Levow, Gina-Anne. Effectiveness of Scalable Monolingual Data and Trigger Words Prompting on Cross-Lingual Emotion Detection Task. Proceedings of the 14th Workshop on Computational Approaches to Subjectivi...
2024
-
[152]
WU \\_ TLAXE at WASSA 2024 Explainability for Cross-Lingual Emotion in Tweets Shared Task 1: Emotion through Translation using T w HIN - BERT and GPT
Davenport, Jon and Ruditsky, Keren and Batra, Anna and Lhawa, Yulha and Levow, Gina-Anne. WU \\_ TLAXE at WASSA 2024 Explainability for Cross-Lingual Emotion in Tweets Shared Task 1: Emotion through Translation using T w HIN - BERT and GPT. Proceedings of the 14th Workshop on ...
2024
-
[153]
Enhancing Cross-Lingual Emotion Detection with Data Augmentation and Token-Label Mapping
Zhang, Jinghui and Zhao, Yuan and Zhang, Siqin and Zhao, Ruijing and Bao, Siyu. Enhancing Cross-Lingual Emotion Detection with Data Augmentation and Token-Label Mapping. Proceedings of the 14th Workshop on Computational Approaches to Subjectivity, Sentiment, \& Social Media An...
2024
-
[154]
Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024
2024
-
[155]
V ar D ial Evaluation Campaign 2024: Commonsense Reasoning in Dialects and Multi-Label Similar Language Identification
Chifu, Adrian-Gabriel and Glava s , Goran and Ionescu, Radu Tudor and Ljube s i \'c , Nikola and Mileti \'c , Aleksandra and Mileti \'c , Filip and Scherrer, Yves and Vuli \'c , Ivan. V ar D ial Evaluation Campaign 2024: Commonsense Reasoning in Dialects and Multi-Label Simila...
2024 doi
-
[156]
What Drives Performance in Multilingual Language Models?
Bagheri Nezhad, Sina and Agrawal, Ameeta. What Drives Performance in Multilingual Language Models?. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.2
2024 doi
-
[157]
Does Whisper Understand S wiss G erman? An Automatic, Qualitative, and Human Evaluation
Dolev, Eyal and Lutz, Clemens and Aepli, No. Does Whisper Understand S wiss G erman? An Automatic, Qualitative, and Human Evaluation. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.3
2024 doi
-
[158]
How Well Do Tweets Represent Sub-Dialects of E gyptian A rabic?
Mohamed Eida, Mai and Nassar, Mayar and Dunn, Jonathan. How Well Do Tweets Represent Sub-Dialects of E gyptian A rabic?. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.4
2024 doi
-
[159]
When Elote, Choclo and Mazorca are not the Same
Espa \ n a-Bonet, Cristina and Bhatt, Ankur and Dutta Chowdhury, Koel and Barr \'o n-Cede \ n o, Alberto. When Elote, Choclo and Mazorca are not the Same. Isomorphism-Based Perspective to the S panish Varieties Divergences. Proceedings of the Eleventh Workshop on NLP for Simil...
2024 doi
-
[160]
Modeling Orthographic Variation in O ccitan ' s Dialects
Hopton, Zachary and Aepli, No. Modeling Orthographic Variation in O ccitan ' s Dialects. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.6
2024 doi
-
[161]
DIALECT - COPA : Extending the Standard Translations of the COPA Causal Commonsense Reasoning Dataset to S outh S lavic Dialects
Ljube s i \'c , Nikola and Galant, Nada and Ben c ina, Sonja and C ibej, Jaka and Milosavljevi \'c , Stefan and Rupnik, Peter and Kuzman, Taja. DIALECT - COPA : Extending the Standard Translations of the COPA Causal Commonsense Reasoning Dataset to S outh S lavic Dialects. Pro...
2024 doi
-
[162]
The Role of Adverbs in Language Variety Identification: The Case of P ortuguese Multi-Word Adverbs
M. The Role of Adverbs in Language Variety Identification: The Case of P ortuguese Multi-Word Adverbs. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.8
2024 doi
-
[163]
N o M usic - The N orwegian Multi-Dialectal Slot and Intent Detection Corpus
M hlum, Petter and Scherrer, Yves. N o M usic - The N orwegian Multi-Dialectal Slot and Intent Detection Corpus. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.9
2024 doi
-
[164]
Understanding Position Bias Effects on Fairness in Social Multi-Document Summarization
Olabisi, Olubusayo and Agrawal, Ameeta. Understanding Position Bias Effects on Fairness in Social Multi-Document Summarization. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.10
2024 doi
-
[165]
Can LLM s Handle Low-Resource Dialects? A Case Study on Translation and Common Sense Reasoning in S ari s
Ondrejov \'a , Vikt \'o ria and S uppa, Marek. Can LLM s Handle Low-Resource Dialects? A Case Study on Translation and Common Sense Reasoning in S ari s. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.186...
2024 doi
-
[166]
Experiments in Multi-Variant Natural Language Processing for N ahuatl
Pugh, Robert and Tyers, Francis. Experiments in Multi-Variant Natural Language Processing for N ahuatl. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.12
2024 doi
-
[167]
Highly Granular Dialect Normalization and Phonological Dialect Translation for L imburgish
Simons, Andreas and De Pascale, Stefano and Franco, Karlien. Highly Granular Dialect Normalization and Phonological Dialect Translation for L imburgish. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.1865...
2024 doi
-
[168]
Multilingual Identification of E nglish Code-Switching
Sterner, Igor. Multilingual Identification of E nglish Code-Switching. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.14
2024 doi
-
[169]
Studying Language Variation Considering the Re-Usability of Modern Theories, Tools and Resources for Annotating Explicit and Implicit Events in Centuries Old Text
Verkijk, Stella and Sommerauer, Pia and Vossen, Piek. Studying Language Variation Considering the Re-Usability of Modern Theories, Tools and Resources for Annotating Explicit and Implicit Events in Centuries Old Text. Proceedings of the Eleventh Workshop on NLP for Similar Lan...
2024 doi
-
[170]
Language Identification of P hilippine Creole S panish: Discriminating C havacano From Related Languages
Vicente, Aileen Joan and Cheng, Charibeth. Language Identification of P hilippine Creole S panish: Discriminating C havacano From Related Languages. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1...
2024 doi
-
[171]
Data-Augmentation-Based Dialectal Adaptation for LLM s
Faisal, Fahim and Anastasopoulos, Antonios. Data-Augmentation-Based Dialectal Adaptation for LLM s. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.17
2024 doi
-
[172]
Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024)
Ljube s i \'c , Nikola and Kuzman, Taja and Rupnik, Peter and Vuli \'c , Ivan and Schmidt, Fabian and Glava s , Goran. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.18
2024 doi
-
[173]
Incorporating Dialect Understanding Into LLM Using RAG and Prompt Engineering Techniques for Causal Commonsense Reasoning
Perak, Benedikt and Beliga, Slobodan and Me s trovi \'c , Ana. Incorporating Dialect Understanding Into LLM Using RAG and Prompt Engineering Techniques for Causal Commonsense Reasoning. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects ...
2024 doi
-
[174]
One-Shot Prompt for Language Variety Identification
Gillin, Nat. One-Shot Prompt for Language Variety Identification. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.20
2024 doi
-
[175]
Improving Multi-Label Classification of Similar Languages by Semantics-Aware Word Embeddings
Ngo, The and Nguyen, Thi Anh and Ha, My and Nguyen, Thi Minh and Le-Hong, Phuong. Improving Multi-Label Classification of Similar Languages by Semantics-Aware Word Embeddings. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2...
2024 doi
-
[176]
B randeis at V ar D ial 2024 DSL - ML Shared Task: Multilingual Models, Simple Baselines and Data Augmentation
S. B randeis at V ar D ial 2024 DSL - ML Shared Task: Multilingual Models, Simple Baselines and Data Augmentation. Proceedings of the Eleventh Workshop on NLP for Similar Languages, Varieties, and Dialects (VarDial 2024). 2024. doi:10.18653/v1/2024.vardial-1.22
2024 doi
-
[177]
Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[178]
A Contemporary News Corpus of U krainian ( CNC - UA ): Compilation, Annotation, Publication
Fischer, Stefan and Haidarzhyi, Kateryna and Knappen, J. A Contemporary News Corpus of U krainian ( CNC - UA ): Compilation, Annotation, Publication. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[179]
Introducing the Djinni Recruitment Dataset: A Corpus of Anonymized CV s and Job Postings
Drushchak, Nazarii and Romanyshyn, Mariana. Introducing the Djinni Recruitment Dataset: A Corpus of Anonymized CV s and Job Postings. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[180]
Creating Parallel Corpora for U krainian: A G erman- U krainian Parallel Corpus ( P ara R ook|| DE - UK )
Shvedova, Maria and Lukashevskyi, Arsenii. Creating Parallel Corpora for U krainian: A G erman- U krainian Parallel Corpus ( P ara R ook|| DE - UK ). Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[181]
Introducing NER - UK 2.0: A Rich Corpus of Named Entities for U krainian
Chaplynskyi, Dmytro and Romanyshyn, Mariana. Introducing NER - UK 2.0: A Rich Corpus of Named Entities for U krainian. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[182]
Instant Messaging Platforms News Multi-Task Classification for Stance, Sentiment, and Discrimination Detection
Ustyianovych, Taras and Barbosa, Denilson. Instant Messaging Platforms News Multi-Task Classification for Stance, Sentiment, and Discrimination Detection. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[183]
Setting up the Data Printer with Improved E nglish to U krainian Machine Translation
Paniv, Yurii and Chaplynskyi, Dmytro and Trynus, Nikita and Kyrylov, Volodymyr. Setting up the Data Printer with Improved E nglish to U krainian Machine Translation. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[184]
Automated Extraction of Hypo-Hypernym Relations for the U krainian W ord N et
Romanyshyn, Nataliia and Chaplynskyi, Dmytro and Romanyshyn, Mariana. Automated Extraction of Hypo-Hypernym Relations for the U krainian W ord N et. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[185]
U krainian Visual Word Sense Disambiguation Benchmark
Laba, Yurii and Mohytych, Yaryna and Rohulia, Ivanna and Kyryleyza, Halyna and Dydyk-Meush, Hanna and Dobosevych, Oles and Hryniv, Rostyslav. U krainian Visual Word Sense Disambiguation Benchmark. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @...
2024
-
[186]
The UNLP 2024 Shared Task on Fine-Tuning Large Language Models for U krainian
Romanyshyn, Mariana and Syvokon, Oleksiy and Kyslyi, Roman. The UNLP 2024 Shared Task on Fine-Tuning Large Language Models for U krainian. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[187]
Fine-Tuning and Retrieval Augmented Generation for Question Answering Using Affordable Large Language Models
Boros, Tiberiu and Chivereanu, Radu and Dumitrescu, Stefan and Purcaru, Octavian. Fine-Tuning and Retrieval Augmented Generation for Question Answering Using Affordable Large Language Models. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC...
2024
-
[188]
From Bytes to Borsch: Fine-Tuning Gemma and Mistral for the U krainian Language Representation
Kiulian, Artur and Polishko, Anton and Khandoga, Mykola and Chubych, Oryna and Connor, Jack and Ravishankar, Raghav and Shirawalmath, Adarsh. From Bytes to Borsch: Fine-Tuning Gemma and Mistral for the U krainian Language Representation. Proceedings of the Third Ukrainian Natu...
2024
-
[189]
Spivavtor: An Instruction Tuned U krainian Text Editing Model
Saini, Aman and Chernodub, Artem and Raheja, Vipul and Kulkarni, Vivek. Spivavtor: An Instruction Tuned U krainian Text Editing Model. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[190]
Eval- UA -tion 1.0: Benchmark for Evaluating U krainian (Large) Language Models
Hamotskyi, Serhii and Levbarg, Anna-Izabella and H. Eval- UA -tion 1.0: Benchmark for Evaluating U krainian (Large) Language Models. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[191]
L i BERT a: Advancing U krainian Language Modeling through Pre-training from Scratch
Haltiuk, Mykola and Smywi \'n ski-Pohl, Aleksander. L i BERT a: Advancing U krainian Language Modeling through Pre-training from Scratch. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[192]
Entity Embellishment Mitigation in LLM s Output with Noisy Synthetic Dataset for Alignment
Galeshchuk, Svitlana. Entity Embellishment Mitigation in LLM s Output with Noisy Synthetic Dataset for Alignment. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[193]
Language-Specific Pruning for Efficient Reduction of Large Language Models
Shamrai, Maksym. Language-Specific Pruning for Efficient Reduction of Large Language Models. Proceedings of the Third Ukrainian Natural Language Processing Workshop (UNLP) @ LREC-COLING 2024. 2024
2024
-
[194]
Proceedings of the Third Workshop on Understanding Implicit and Underspecified Language. 2024
2024
-
[195]
Taking Action Towards Graceful Interaction: The Effects of Performing Actions on Modelling Policies for Instruction Clarification Requests
Madureira, Brielen and Schlangen, David. Taking Action Towards Graceful Interaction: The Effects of Performing Actions on Modelling Policies for Instruction Clarification Requests. Proceedings of the Third Workshop on Understanding Implicit and Underspecified Language. 2024
2024
-
[196]
More Labels or Cases? Assessing Label Variation in Natural Language Inference
Gruber, Cornelia and Hechinger, Katharina and Assenmacher, Matthias and Kauermann, G. More Labels or Cases? Assessing Label Variation in Natural Language Inference. Proceedings of the Third Workshop on Understanding Implicit and Underspecified Language. 2024
2024
-
[197]
Resolving Transcription Ambiguity in S panish: A Hybrid Acoustic-Lexical System for Punctuation Restoration
Zhu, Xiliang and Chang, Chia-Tien and Gardiner, Shayna and Rossouw, David and Robertson, Jonas. Resolving Transcription Ambiguity in S panish: A Hybrid Acoustic-Lexical System for Punctuation Restoration. Proceedings of the Third Workshop on Understanding Implicit and Underspe...
2024
-
[198]
Assessing the Significance of Encoded Information in Contextualized Representations to Word Sense Disambiguation
Yavas, Deniz Ekin. Assessing the Significance of Encoded Information in Contextualized Representations to Word Sense Disambiguation. Proceedings of the Third Workshop on Understanding Implicit and Underspecified Language. 2024
2024
-
[199]
Below the Sea (with the Sharks): Probing Textual Features of Implicit Sentiment in a Literary Case-study
Bizzoni, Yuri and Feldkamp, Pascale. Below the Sea (with the Sharks): Probing Textual Features of Implicit Sentiment in a Literary Case-study. Proceedings of the Third Workshop on Understanding Implicit and Underspecified Language. 2024
2024
-
[200]
Exposing propaganda: an analysis of stylistic cues comparing human annotations and machine classification
Faye, G \'e raud and Icard, Benjamin and Casanova, Morgane and Chanson, Julien and Maine, Fran c ois and Bancilhon, Fran c ois and Gadek, Guillaume and Gravier, Guillaume and \'E gr \'e , Paul. Exposing propaganda: an analysis of stylistic cues comparing human annotations and ...
2024
-
[201]
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
Peng, Siyao and Sun, Zihang and Loftus, Sebastian and Plank, Barbara. Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations. Proceedings of the Third Workshop on Understanding Implicit and Underspecified Language. 2024
2024
-
[202]
Colour Me Uncertain: Representing Vagueness with Probabilistic Semantics
Chun Cheung, Kin and Emerson, Guy. Colour Me Uncertain: Representing Vagueness with Probabilistic Semantics. Proceedings of the Third Workshop on Understanding Implicit and Underspecified Language. 2024
2024
-
[203]
Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[204]
Calibration-Tuning: Teaching Large Language Models to Know What They Don ' t Know
Kapoor, Sanyam and Gruver, Nate and Roberts, Manley and Pal, Arka and Dooley, Samuel and Goldblum, Micah and Wilson, Andrew. Calibration-Tuning: Teaching Large Language Models to Know What They Don ' t Know. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP...
2024
-
[205]
Context Tuning for Retrieval Augmented Generation
Anantha, Raviteja and Vodianik, Danil. Context Tuning for Retrieval Augmented Generation. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[206]
Optimizing Relation Extraction in Medical Texts through Active Learning: A Comparative Analysis of Trade-offs
Liang, Siting and Valdunciel S \'a nchez, Pablo and Sonntag, Daniel. Optimizing Relation Extraction in Medical Texts through Active Learning: A Comparative Analysis of Trade-offs. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[207]
Linguistic Obfuscation Attacks and Large Language Model Uncertainty
Steindl, Sebastian and Sch. Linguistic Obfuscation Attacks and Large Language Model Uncertainty. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[208]
Aligning Uncertainty: Leveraging LLM s to Analyze Uncertainty Transfer in Text Summarization
Kolagar, Zahra and Zarcone, Alessandra. Aligning Uncertainty: Leveraging LLM s to Analyze Uncertainty Transfer in Text Summarization. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[209]
How Does Beam Search improve Span-Level Confidence Estimation in Generative Sequence Labeling?
Hashimoto, Kazuma and Naim, Iftekhar and Raman, Karthik. How Does Beam Search improve Span-Level Confidence Estimation in Generative Sequence Labeling?. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[210]
Efficiently Acquiring Human Feedback with B ayesian Deep Learning
Fang, Haishuo and Gor, Jeet and Simpson, Edwin. Efficiently Acquiring Human Feedback with B ayesian Deep Learning. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[211]
Order Effects in Annotation Tasks: Further Evidence of Annotation Sensitivity
Beck, Jacob and Eckman, Stephanie and Ma, Bolei and Chew, Rob and Kreuter, Frauke. Order Effects in Annotation Tasks: Further Evidence of Annotation Sensitivity. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[212]
The Effect of Generalisation on the Inadequacy of the Mode
Eikema, Bryan. The Effect of Generalisation on the Inadequacy of the Mode. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[213]
Uncertainty Resolution in Misinformation Detection
Orlovskiy, Yury and Thibault, Camille and Imouza, Anne and Godbout, Jean-Fran c ois and Rabbany, Reihaneh and Pelrine, Kellin. Uncertainty Resolution in Misinformation Detection. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[214]
Don ' t Blame the Data, Blame the Model: Understanding Noise and Bias When Learning from Subjective Annotations
Anand, Abhishek and Mokhberian, Negar and Kumar, Prathyusha and Saha, Anweasha and He, Zihao and Rao, Ashwin and Morstatter, Fred and Lerman, Kristina. Don ' t Blame the Data, Blame the Model: Understanding Noise and Bias When Learning from Subjective Annotations. Proceedings ...
2024
-
[215]
Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation
Rivera, Mauricio and Godbout, Jean-Fran c ois and Rabbany, Reihaneh and Pelrine, Kellin. Combining Confidence Elicitation and Sample-based Methods for Uncertainty Quantification in Misinformation Mitigation. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP...
2024
-
[216]
Linguistically Communicating Uncertainty in Patient-Facing Risk Prediction Models
Sivaprasad, Adarsa and Reiter, Ehud. Linguistically Communicating Uncertainty in Patient-Facing Risk Prediction Models. Proceedings of the 1st Workshop on Uncertainty-Aware NLP (UncertaiNLP 2024). 2024
2024
-
[217]
Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024
2024
-
[218]
Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text
Adilazuarda, Muhammad. Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.1
2024 doi
-
[219]
Automated Adversarial Discovery for Safety Classifiers
Lal, Yash Kumar and Lahoti, Preethi and Sinha, Aradhana and Qin, Yao and Balashankar, Ananth. Automated Adversarial Discovery for Safety Classifiers. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.2
2024 doi
-
[220]
F air B elief - Assessing Harmful Beliefs in Language Models
Setzu, Mattia and Marchiori Manerba, Marta and Minervini, Pasquale and Nozza, Debora. F air B elief - Assessing Harmful Beliefs in Language Models. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.3
2024 doi
-
[221]
The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification
Bui, Minh Duc and Von Der Wense, Katharina. The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.4
2024 doi
-
[222]
When XGB oost Outperforms GPT -4 on Text Classification: A Case Study
Bohacek, Matyas and Bravansky, Michal. When XGB oost Outperforms GPT -4 on Text Classification: A Case Study. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.5
2024 doi
-
[223]
Towards Healthy AI : Large Language Models Need Therapists Too
Lin, Baihan and Bouneffouf, Djallel and Cecchi, Guillermo and Varshney, Kush. Towards Healthy AI : Large Language Models Need Therapists Too. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.6
2024 doi
-
[224]
Exploring Causal Mechanisms for Machine Text Detection Methods
Yoo, Kiyoon and Ahn, Wonhyuk and Song, Yeji and Kwak, Nojun. Exploring Causal Mechanisms for Machine Text Detection Methods. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.7
2024 doi
-
[225]
F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment
Rashad, Mohamed and Zahran, Ahmed and Amin, Abanoub and Abdelaal, Amr and Altantawy, Mohamed. F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (Trus...
2024 doi
-
[226]
Cross-Task Defense: Instruction-Tuning LLM s for Content Safety
Fu, Yu and Xiao, Wen and Chen, Jia and Li, Jiachen and Papalexakis, Evangelos and Chien, Aichi and Dong, Yue. Cross-Task Defense: Instruction-Tuning LLM s for Content Safety. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:...
2024 doi
-
[227]
On the Interplay between Fairness and Explainability
Brandl, Stephanie and Bugliarello, Emanuele and Chalkidis, Ilias. On the Interplay between Fairness and Explainability. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.10
2024 doi
-
[228]
Holistic Evaluation of Large Language Models: Assessing Robustness, Accuracy, and Toxicity for Real-World Applications
Cecchini, David and Nazir, Arshaan and Chakravarthy, Kalyan and Kocaman, Veysel. Holistic Evaluation of Large Language Models: Assessing Robustness, Accuracy, and Toxicity for Real-World Applications. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (...
2024 doi
-
[229]
HGOT : Hierarchical Graph of Thoughts for Retrieval-Augmented In-Context Learning in Factuality Evaluation
Fang, Yihao and Thomas, Stephen and Zhu, Xiaodan. HGOT : Hierarchical Graph of Thoughts for Retrieval-Augmented In-Context Learning in Factuality Evaluation. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024....
2024 doi
-
[230]
Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models
Groot, Tobias and Valdenegro - Toro, Matias. Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.13
2024 doi
-
[231]
Tweak to Trust: Assessing the Reliability of Summarization Metrics in Contact Centers via Perturbed Summaries
Patel, Kevin and Agrawal, Suraj and Kumar, Ayush. Tweak to Trust: Assessing the Reliability of Summarization Metrics in Contact Centers via Perturbed Summaries. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/20...
2024 doi
-
[232]
Flatness-Aware Gradient Descent for Safe Conversational AI
Khalatbari, Leila and Hosseini, Saeid and Sameti, Hossein and Fung, Pascale. Flatness-Aware Gradient Descent for Safe Conversational AI. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.15
2024 doi
-
[233]
Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations
Cao, Lele and Buchner, Valentin and Senane, Zineb and Yang, Fangkai. Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.tr...
2024 doi
-
[234]
Semantic-Preserving Adversarial Example Attack against BERT
Gao, Chongyang and Gu, Kang and Vosoughi, Soroush and Mehnaz, Shagufta. Semantic-Preserving Adversarial Example Attack against BERT. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.17
2024 doi
-
[235]
Sandwich attack: Multi-language Mixture Adaptive Attack on LLM s
Upadhayay, Bibek and Behzadan, Vahid. Sandwich attack: Multi-language Mixture Adaptive Attack on LLM s. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustNLP 2024). 2024. doi:10.18653/v1/2024.trustnlp-1.18
2024 doi
-
[236]
Masking Latent Gender Knowledge for Debiasing Image Captioning
Yang, Fan and Ghosh, Shalini and Barut, Emre and Qin, Kechen and Wanigasekara, Prashan and Su, Chengwei and Ruan, Weitong and Gupta, Rahul. Masking Latent Gender Knowledge for Debiasing Image Captioning. Proceedings of the 4th Workshop on Trustworthy Natural Language Processin...
2024 doi
-
[237]
BELIEVE : Belief-Enhanced Instruction Generation and Augmentation for Zero-Shot Bias Mitigation
Bauer, Lisa and Mehrabi, Ninareh and Goyal, Palash and Chang, Kai-Wei and Galstyan, Aram and Gupta, Rahul. BELIEVE : Belief-Enhanced Instruction Generation and Augmentation for Zero-Shot Bias Mitigation. Proceedings of the 4th Workshop on Trustworthy Natural Language Processin...
2024 doi
-
[238]
Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models
Zarharan, Majid and Wullschleger, Pascal and Behkam Kia, Babak and Pilehvar, Mohammad Taher and Foster, Jennifer. Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models. Proceedings of the 4th Workshop on Trustworthy Natural Language Processing (TrustN...
2024 doi
-
[239]
Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[240]
The Constant in HATE : Toxicity in R eddit across Topics and Languages
Tufa, Wondimagegnhue Tsegaye and Markov, Ilia and Vossen, Piek T.J.M. The Constant in HATE : Toxicity in R eddit across Topics and Languages. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[241]
A Federated Learning Approach to Privacy Preserving Offensive Language Identification
Zampieri, Marcos and Premasiri, Damith and Ranasinghe, Tharindu. A Federated Learning Approach to Privacy Preserving Offensive Language Identification. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[242]
CLTL @ H arm P ot- ID : Leveraging Transformer Models for Detecting Offline Harm Potential and Its Targets in Low-Resource Languages
Wang, Yeshan and Markov, Ilia. CLTL @ H arm P ot- ID : Leveraging Transformer Models for Detecting Offline Harm Potential and Its Targets in Low-Resource Languages. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[243]
NJUST - KMG at TRAC -2024 Tasks 1 and 2: Offline Harm Potential Identification
Wang, Jingyuan and Depp, Jack and Yang, Yang. NJUST - KMG at TRAC -2024 Tasks 1 and 2: Offline Harm Potential Identification. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[244]
and Jha, Soumya Sangam and Rao, Vartika T
H C, Anagha and Krishna, Saatvik M. and Jha, Soumya Sangam and Rao, Vartika T. and M, Anand Kumar. S calar L ab@ TRAC 2024: Exploring Machine Learning Techniques for Identifying Potential Offline Harm in Multilingual Commentaries. Proceedings of the Fourth Workshop on Threat, ...
2024
-
[245]
LLM -Based Synthetic Datasets: Applications and Limitations in Toxicity Detection
Kruschwitz, Udo and Schmidhuber, Maximilian. LLM -Based Synthetic Datasets: Applications and Limitations in Toxicity Detection. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[246]
Using Sarcasm to Improve Cyberbullying Detection
Guo, Xiaoyu and Gauch, Susan. Using Sarcasm to Improve Cyberbullying Detection. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[247]
Analyzing Offensive Language and Hate Speech in Political Discourse: A Case Study of G erman Politicians
Weissenbacher, Maximilian and Kruschwitz, Udo. Analyzing Offensive Language and Hate Speech in Political Discourse: A Case Study of G erman Politicians. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[248]
Ice and Fire: Dataset on Sentiment, Emotions, Toxicity, Sarcasm, Hate speech, Sympathy and More in I celandic Blog Comments
Fri riksd \'o ttir, Steinunn Rut and Simonsen, Annika and \'A smundsson, Atli Sn r and Fri j \'o nsd \'o ttir, Gu r \'u n Lilja and Ingason, Anton Karl and Sn bjarnarson, V \'e steinn and Einarsson, Hafsteinn. Ice and Fire: Dataset on Sentiment, Emotions, Toxicity, Sarcasm, Ha...
2024
-
[249]
Detecting Hate Speech in A mharic Using Multimodal Analysis of Social Media Memes
Jigar, Melese Ayichlie and Ayele, Abinew Ali and Yimam, Seid Muhie and Biemann, Chris. Detecting Hate Speech in A mharic Using Multimodal Analysis of Social Media Memes. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[250]
Content Moderation in Online Platforms: A Study of Annotation Methods for Inappropriate Language
Barbarestani, Baran and Maks, Isa and Vossen, Piek T.J.M. Content Moderation in Online Platforms: A Study of Annotation Methods for Inappropriate Language. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[251]
F rench T oxicity P rompts: a Large Benchmark for Evaluating and Mitigating Toxicity in F rench Texts
Brun, Caroline and Nikoulina, Vassilina. F rench T oxicity P rompts: a Large Benchmark for Evaluating and Mitigating Toxicity in F rench Texts. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[252]
Studying Reactions to Stereotypes in Teenagers: an Annotated I talian Dataset
Chierchiello, Elisa and Bourgeade, Tom and Ricci, Giacomo and Bosco, Cristina and D ' Errico, Francesca. Studying Reactions to Stereotypes in Teenagers: an Annotated I talian Dataset. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[253]
Offensiveness, Hate, Emotion and GPT : Benchmarking GPT 3.5 and GPT 4 as Classifiers on T witter-specific Datasets
Bauer, Nikolaj and Preisig, Moritz and Volk, Martin. Offensiveness, Hate, Emotion and GPT : Benchmarking GPT 3.5 and GPT 4 as Classifiers on T witter-specific Datasets. Proceedings of the Fourth Workshop on Threat, Aggression \& Cyberbullying @ LREC-COLING-2024. 2024
2024
-
[254]
D o D o Learning: Domain-Demographic Transfer in Language Models for Detecting Abuse Targeted at Public Figures
Williams, Angus Redlarski and Kirk, Hannah Rose and Burke-Moore, Liam and Chung, Yi-Ling and Debono, Ivan and Johansson, Pica and Stevens, Francesca and Bright, Jonathan and Hale, Scott. D o D o Learning: Domain-Demographic Transfer in Language Models for Detecting Abuse Targe...
2024
-
[255]
Empowering Users and Mitigating Harm: Leveraging Nudging Principles to Enhance Social Media Safety
Donabauer, Gregor and Theophilou, Emily and Lomonaco, Francesco and Bursic, Sathya and Taibi, Davide and Hern \'a ndez-Leo, Davinia and Kruschwitz, Udo and Ognibene, Dimitri. Empowering Users and Mitigating Harm: Leveraging Nudging Principles to Enhance Social Media Safety. Pr...
2024
-
[256]
Exploring Boundaries and Intensities in Offensive and Hate Speech: Unveiling the Complex Spectrum of Social Media Discourse
Ayele, Abinew Ali and Jalew, Esubalew Alemneh and Ali, Adem Chanie and Yimam, Seid Muhie and Biemann, Chris. Exploring Boundaries and Intensities in Offensive and Hate Speech: Unveiling the Complex Spectrum of Social Media Discourse. Proceedings of the Fourth Workshop on Threa...
2024
-
[257]
Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[258]
Learning Human Action Representations from Temporal Context in Lifestyle Vlogs
Ignat, Oana and Castro, Santiago and Li, Weiji and Mihalcea, Rada. Learning Human Action Representations from Temporal Context in Lifestyle Vlogs. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[259]
C on G ra T : Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings
Brannon, William and Kang, Wonjune and Fulay, Suyash and Jiang, Hang and Roy, Brandon and Roy, Deb and Kabbara, Jad. C on G ra T : Self-Supervised Contrastive Pretraining for Joint Graph and Text Embeddings. Proceedings of TextGraphs-17: Graph-based Methods for Natural Languag...
2024
-
[260]
Uniform Meaning Representation Parsing as a Pipelined Approach
Chun, Jayeol and Xue, Nianwen. Uniform Meaning Representation Parsing as a Pipelined Approach. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[261]
Financial Product Ontology Population with Large Language Models
Saetia, Chanatip and Phruetthiset, Jiratha and Chalothorn, Tawunrat and Lertsutthiwong, Monchai and Taerungruang, Supawat and Buabthong, Pakpoom. Financial Product Ontology Population with Large Language Models. Proceedings of TextGraphs-17: Graph-based Methods for Natural Lan...
2024
-
[262]
Prompt Me One More Time: A Two-Step Knowledge Extraction Pipeline with Ontology-Based Verification
Chepurova, Alla and Kuratov, Yuri and Bulatov, Aydar and Burtsev, Mikhail. Prompt Me One More Time: A Two-Step Knowledge Extraction Pipeline with Ontology-Based Verification. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[263]
Towards Understanding Attention-based Reasoning through Graph Structures in Medical Codes Classification
Goldstein, Noon and Amin, Saadullah and Neumann, G. Towards Understanding Attention-based Reasoning through Graph Structures in Medical Codes Classification. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[264]
Leveraging Graph Structures to Detect Hallucinations in Large Language Models
Nonkes, Noa and Agaronian, Sergei and Kanoulas, Evangelos and Petcu, Roxana. Leveraging Graph Structures to Detect Hallucinations in Large Language Models. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[265]
Semantic Graphs for Syntactic Simplification: A Revisit from the Age of LLM
Yao, Peiran and Guzhva, Kostyantyn and Barbosa, Denilson. Semantic Graphs for Syntactic Simplification: A Revisit from the Age of LLM. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[266]
T ext G raphs 2024 Shared Task on Text-Graph Representations for Knowledge Graph Question Answering
Sakhovskiy, Andrey and Salnikov, Mikhail and Nikishina, Irina and Usmanova, Aida and Kraft, Angelie and M. T ext G raphs 2024 Shared Task on Text-Graph Representations for Knowledge Graph Question Answering. Proceedings of TextGraphs-17: Graph-based Methods for Natural Languag...
2024
-
[267]
nlp \\_ enjoyers at T ext G raphs-17 Shared Task: Text-Graph Representations for Knowledge Graph Question Answering using all- MPN et
Kurdiukov, Nikita and Zinkovich, Viktoriia and Karpukhin, Sergey and Tikhomirov, Pavel. nlp \\_ enjoyers at T ext G raphs-17 Shared Task: Text-Graph Representations for Knowledge Graph Question Answering using all- MPN et. Proceedings of TextGraphs-17: Graph-based Methods for ...
2024
-
[268]
HW - TSC at T ext G raphs-17 Shared Task: Enhancing Inference Capabilities of LLM s with Knowledge Graphs
Tang, Wei and Qiao, Xiaosong and Zhao, Xiaofeng and Zhang, Min and Su, Chang and Li, Yuang and Li, Yinglu and Liu, Yilun and Yao, Feiyu and Tao, Shimin and Yang, Hao and Xianghui, He. HW - TSC at T ext G raphs-17 Shared Task: Enhancing Inference Capabilities of LLM s with Know...
2024
-
[269]
TIGFORMER at T ext G raphs-17 Shared Task: A Late Interaction Method for text and Graph Representations in KBQA Classification Task
Rakesh, Mayank and Saikia, Parikshit and Shrivastava, Saket. TIGFORMER at T ext G raphs-17 Shared Task: A Late Interaction Method for text and Graph Representations in KBQA Classification Task. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[270]
NLP eople at T ext G raphs-17 Shared Task: Chain of Thought Questioning to Elicit Decompositional Reasoning
Moses, Movina and Kuruvanthodi, Vishnudev and Elkaref, Mohab and Tanaka, Shinnosuke and Barry, James and Mel, Geeth and Watson, Campbell. NLP eople at T ext G raphs-17 Shared Task: Chain of Thought Questioning to Elicit Decompositional Reasoning. Proceedings of TextGraphs-17: ...
2024
-
[271]
Skoltech at T ext G raphs-17 Shared Task: Finding GPT -4 Prompting Strategies for Multiple Choice Questions
Lysyuk, Maria and Braslavski, Pavel. Skoltech at T ext G raphs-17 Shared Task: Finding GPT -4 Prompting Strategies for Multiple Choice Questions. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[272]
J elly B ell at T ext G raphs-17 Shared Task: Fusing Large Language Models with External Knowledge for Enhanced Question Answering
Belikova, Julia and Beliakin, Evegeniy and Konovalov, Vasily. J elly B ell at T ext G raphs-17 Shared Task: Fusing Large Language Models with External Knowledge for Enhanced Question Answering. Proceedings of TextGraphs-17: Graph-based Methods for Natural Language Processing. 2024
2024
-
[273]
Proceedings of the 1st Worskhop on Towards Ethical and Inclusive Conversational AI: Language Attitudes, Linguistic Diversity, and Language Rights (TEICAI 2024). 2024
2024
-
[274]
How Do Conversational Agents in Healthcare Impact on Patient Agency?
Denecke, Kerstin. How Do Conversational Agents in Healthcare Impact on Patient Agency?. Proceedings of the 1st Worskhop on Towards Ethical and Inclusive Conversational AI: Language Attitudes, Linguistic Diversity, and Language Rights (TEICAI 2024). 2024
2024
-
[275]
Why academia should cut back general enthusiasm about CA s
Giulimondi, Alessia. Why academia should cut back general enthusiasm about CA s. Proceedings of the 1st Worskhop on Towards Ethical and Inclusive Conversational AI: Language Attitudes, Linguistic Diversity, and Language Rights (TEICAI 2024). 2024
2024
-
[276]
Bridging the Language Gap: Integrating Language Variations into Conversational AI Agents for Enhanced User Engagement
Amadeus, Marcellus and Homeli da Silva, Jose Roberto and Pessoa Rocha, Joao Victor. Bridging the Language Gap: Integrating Language Variations into Conversational AI Agents for Enhanced User Engagement. Proceedings of the 1st Worskhop on Towards Ethical and Inclusive Conversat...
2024
-
[277]
Socio-cultural adapted chatbots: Harnessing Knowledge Graphs and Large Language Models for enhanced context awarenes
Camboim de S \'a , Jader and Anastasiou, Dimitra and Da Silveira, Marcos and Pruski, C \'e dric. Socio-cultural adapted chatbots: Harnessing Knowledge Graphs and Large Language Models for enhanced context awarenes. Proceedings of the 1st Worskhop on Towards Ethical and Inclusi...
2024
-
[278]
How should Conversational Agent systems respond to sexual harassment?
De Grazia, Laura and Peir \'o Lilja, Alex and Farr \'u s Cabeceran, Mireia and Taul \'e , Mariona. How should Conversational Agent systems respond to sexual harassment?. Proceedings of the 1st Worskhop on Towards Ethical and Inclusive Conversational AI: Language Attitudes, Lin...
2024
-
[279]
Non-Referential Functions of Language in Social Agents: The Case of Social Proximity
H. Non-Referential Functions of Language in Social Agents: The Case of Social Proximity. Proceedings of the 1st Worskhop on Towards Ethical and Inclusive Conversational AI: Language Attitudes, Linguistic Diversity, and Language Rights (TEICAI 2024). 2024
2024
-
[280]
Making a Long Story Short in Conversation Modeling
Tao, Yufei and Mines, Tiernan and Agrawal, Ameeta. Making a Long Story Short in Conversation Modeling. Proceedings of the 1st Worskhop on Towards Ethical and Inclusive Conversational AI: Language Attitudes, Linguistic Diversity, and Language Rights (TEICAI 2024). 2024
2024
-
[281]
Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[282]
Documenting the Unwritten Curriculum of Student Research
Wilson, Shomir. Documenting the Unwritten Curriculum of Student Research. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[283]
Example-Driven Course Slides on Natural Language Processing Concepts
Parde, Natalie. Example-Driven Course Slides on Natural Language Processing Concepts. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[284]
Industry vs Academia: Running a Course on Transformers in Two Setups
Nikishina, Irina and Tikhonova, Maria and Chekalina, Viktoriia and Zaytsev, Alexey and Vazhentsev, Artem and Panchenko, Alexander. Industry vs Academia: Running a Course on Transformers in Two Setups. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[285]
Striking a Balance between Classical and Deep Learning Approaches in Natural Language Processing Pedagogy
Joshi, Aditya and Renzella, Jake and Bhattacharyya, Pushpak and Jha, Saurav and Zhang, Xiangyu. Striking a Balance between Classical and Deep Learning Approaches in Natural Language Processing Pedagogy. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[286]
Co-Creational Teaching of Natural Language Processing
McCrae, John. Co-Creational Teaching of Natural Language Processing. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[287]
Collaborative Development of Modular Open Source Educational Resources for Natural Language Processing
A. Collaborative Development of Modular Open Source Educational Resources for Natural Language Processing. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[288]
From Hate Speech to Societal Empowerment: A Pedagogical Journey Through Computational Thinking and NLP for High School Students
Cignarella, Alessandra Teresa and Chierchiello, Elisa and Ferrando, Chiara and Frenda, Simona and Lo, Soda Marem and Marra, Andrea. From Hate Speech to Societal Empowerment: A Pedagogical Journey Through Computational Thinking and NLP for High School Students. Proceedings of t...
2024
-
[289]
Tightly Coupled Worksheets and Homework Assignments for NLP
Biester, Laura and Wu, Winston. Tightly Coupled Worksheets and Homework Assignments for NLP. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[290]
Teaching LLM s at C harles U niversity: Assignments and Activities
Helcl, Jind r ich and Kasner, Zden e k and Du s ek, Ond r ej and Limisiewicz, Tomasz and Mach \'a c ek, Dominik and Musil, Tom \'a s and Libovick \'y , Jind r ich. Teaching LLM s at C harles U niversity: Assignments and Activities. Proceedings of the Sixth Workshop on Teaching...
2024
-
[291]
Empowering the Future with Multilinguality and Language Diversity
Lee, En-Shiun and Uemura, Kosei and Wasti, Syed and Shipton, Mason. Empowering the Future with Multilinguality and Language Diversity. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[292]
A Course Shared Task on Evaluating LLM Output for Clinical Questions
Hou, Yufang and Tran, Thy and Vu, Doan and Cao, Yiwen and Li, Kai and Rohde, Lukas and Gurevych, Iryna. A Course Shared Task on Evaluating LLM Output for Clinical Questions. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[293]
A Prompting Assignment for Exploring Pretrained LLM s
Anderson, Carolyn. A Prompting Assignment for Exploring Pretrained LLM s. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[294]
Teaching Natural Language Processing in Law School
Braun, Daniel. Teaching Natural Language Processing in Law School. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[295]
Exploring Language Representation through a Resource Inventory Project
Anderson, Carolyn. Exploring Language Representation through a Resource Inventory Project. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[296]
BELT : Building Endangered Language Technology
Ginn, Michael and Saavedra-Beltr \'a n, David and Robayo, Camilo and Palmer, Alexis. BELT : Building Endangered Language Technology. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[297]
Training an NLP Scholar at a Small Liberal Arts College: A Backwards Designed Course Proposal
Prasad, Grusha and Davis, Forrest. Training an NLP Scholar at a Small Liberal Arts College: A Backwards Designed Course Proposal. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[298]
An Interactive Toolkit for Approachable NLP
Brown, AriaRay and Steuer, Julius and Mosbach, Marius and Klakow, Dietrich. An Interactive Toolkit for Approachable NLP. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[299]
Occam ' s Razor and Bender and Koller ' s Octopus
Guerzhoy, Michael. Occam ' s Razor and Bender and Koller ' s Octopus. Proceedings of the Sixth Workshop on Teaching NLP. 2024
2024
-
[300]
Proceedings of the Second International Workshop Towards Digital Language Equality (TDLE): Focusing on Sustainability @ LREC-COLING 2024. 2024
2024
Reviewed June 28, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.