Pith. sign in

REVIEW 3 major objections 5 minor 66 references

MetaphorShare: A Dynamic Collaborative Repository of Open Metaphor Datasets

T0 review · 3 major / 5 minor · reviewed 2026-08-12 · deepseek-v4-flash

Pith's one-line read A new web repository, MetaphorShare, unifies 25 open metaphor datasets under one searchable and re-usable format, with upload, download, search, and labeling tools built in.

desk verdict A useful, honestly scoped repository paper; the main soft spot is the unproven claim that the unified format preserves all original dataset information. read the letter →

arxiv 2411.18260 v3 pith:NSBU6P63 submitted 2024-11-27 cs.CL

classification cs.CL
keywords metaphordatasetsopenrepositoryidentificationdatasetformatunificationMIPVUannotationtoolcross-datasetevaluationElasticsearch
verification ladder T0 review T1 audit T2 compute T3 formal

The pith

A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.

The reading

MetaphorShare is a live website that collects metaphor-labelled corpora from different research traditions and makes them available through a common interface. The paper argues that a single, minimally constrained CSV format can hold datasets that differ in annotation units, context length, and additional variables, and that the repository preserves the original information. If this works, researchers in linguistics and NLP no longer have to hunt for scattered or unpublished datasets or reformat them by hand, and metaphor-identification systems can be trained and evaluated across many datasets at once. The platform currently hosts 25 English datasets and includes an online annotation tool whose output feeds directly into the repository.

What carries the argument

The load-bearing device is the unified CSV format: each line must contain a tagged_text column with at least one tagged expression, using the tag set <m>, <l>, <t>, <a>, <u>, and optional extra columns for sentence index, reference, part of speech, target, source/target domain, metaphoricity scores, or free fields. This format is what lets the repository compare, search, and download heterogeneous datasets through one database schema. Around it sits the website architecture: a FastAPI backend, a PostgreSQL store for dataset metadata, an Elasticsearch index of 'potentially metaphoric expressions' that powers exact and fuzzy multilingual search, and a browser annotation tool that emits the same CSV format.

What would settle it

Download a converted dataset from the repository and compare it against the authors' original release; any instance whose document-level context, multiple labels per span, or continuous score (e.g., metaphoricity rating) is missing or altered would falsify the claim that the unified format preserves original information.

Watch

Extended reading notes

Core claim

The paper presents MetaphorShare as a functioning open repository, currently holding 25 English metaphor datasets, organised around four functionalities: upload, download, search, and label. Its central claim is that a minimally constrained CSV format with five predefined tags — <m> for metaphoric, <l> for literal, <t> for target cue, <a> for anomalous, and <u> for free user-defined tagging — plus optional named columns for continuous or categorical variables, is flexible enough to integrate datasets spanning psycholinguistic word norms, NLP binary classification, multi-word idiom detection, and MIPVU-based full-corpus annotation, without destroying the information the original datasets encode. The paper supports this by converting twelve representative datasets, describing the Elasticsearch-based search and PostgreSQL-backed upload pipeline, and running a cross-dataset RoBERTa metaphor-identification experiment to show that the repository makes multi-dataset experimentation straightforward.

Load-bearing premise

The unified CSV format with its five tags and optional columns preserves all the information contained in the original metaphor datasets, including multiple annotations per expression, document-level context, and continuous variables such as metaphoricity scores.

Editorial extensions

If this is right

  • New datasets formatted as tagged CSV can be uploaded by any user, and after automatic format validation plus manual license review they appear in the catalog and search index.
  • A cross-dataset evaluation on ten training sets shows that models fine-tuned on one metaphor dataset transfer to others to varying degrees, with short, syntactically constrained sets (J&C, GUT) generalising unexpectedly well.
  • The search page allows filtering by dataset, language, and label, with exact and fuzzy matching on tagged expressions or full text, and any result set can be downloaded as CSV.
  • The annotation tool produces labels directly compatible with the repository format, lowering the barrier for annotators without programming experience.

Reading between the lines

Editorial extensions of the paper, not claims the author makes directly.

  • If the repository grows to include non-English datasets as planned, the same CSV format and Elasticsearch multilingual search could make cross-lingual metaphor studies directly comparable, something the current English-only collection only gestures at.
  • The five-tag system may end up functioning as a de facto interchange standard for metaphor annotation, analogous to what CoNLL formats did for syntactic annotation; that would be a larger consequence than the paper itself claims.
  • A natural stress test would be to upload datasets with overlapping but distinct annotation guidelines (MIPVU vs. idiom detection) and measure whether the unified search index supports reliable label-consistent retrieval; the paper does not run that test.
Share X Bluesky LinkedIn Reddit HN

Editorial analysis

A structured set of objections, weighed in public.

Desk editor's note, referee report, and a circularity audit.

Referee Report

3 major / 5 minor

Summary. The paper introduces MetaphorShare, a web-based repository for metaphor datasets, and describes its four main functionalities: upload, download, search, and label. The authors argue that the platform unifies heterogeneous metaphor annotation formats into a single CSV schema with tags <m>, <l>, <t>, <a>, and <u>, while preserving the information in the original datasets. The paper reports that 25 English datasets are currently integrated, describes the system architecture (FastAPI backend, Elasticsearch, PostgreSQL, React frontend), and presents a case study in which RoBERTa models are fine-tuned on twelve of the datasets for cross-dataset metaphor identification. The central claim is that MetaphorShare is a functioning, open, collaborative resource that makes metaphor datasets more accessible and interoperable.

Significance. If the platform works as described, it addresses a genuine gap: many metaphor-labelled resources are scattered, inconsistently formatted, or not known outside small communities. The paper provides a concrete system description, public access, and an illustration of how the repository can support comparative NLP experiments. The strengths are the clear architecture description, the inclusion of diverse datasets, and the fact that the platform is openly accessible with a unified searchable format. However, the paper's own evidence for the preservation of original annotation information is incomplete, and the cross-dataset evaluation is too under-specified to support its comparative claims. These issues are fixable and do not undermine the basic utility of the resource, but they need to be addressed before the paper is ready for publication.

major comments (3)
  1. [Section 3, 'Unified input format'] The paper states in Section 1 that the repository preserves 'the information encoded in the original datasets,' but Section 3 does not provide a mapping from the original annotation schemes to the five-tag CSV format. In particular, MIPVU-derived resources label every token and use categories beyond a binary metaphor/literal distinction, and several datasets include continuous variables such as metaphoricity, novelty, or emotion ratings plus document-level context. The free <u> tag and open columns are flexible, but no example or round-trip check demonstrates that these survive conversion; the two VUAC versions (multiple tags per sentence vs. one tag with duplicated sentences) suggest that choices are being made that could lose information. I request an explicit conversion protocol per dataset type, or a statement of which information is normalised away, and ideally a reconstruction test showing that the original files can be recovered from the stored records.
  2. [Section 3 and Table 4] Section 3 says the twelve illustrative datasets have open licenses, but Table 4 lists PVC as having 'no license.' This contradicts the paper's claim of an open repository and raises legal questions about redistribution. The authors should clarify whether PVC is actually included in MetaphorShare and under what terms, or replace it with a properly licensed dataset in the list of representative examples.
  3. [Section 5, 'Experimental setting' and 'Results'] The cross-dataset evaluation is presented as a case study, but it reports a single F1 score per condition without stating random seeds, number of runs, error bars, or significance tests. As a result, statements such as 'A few datasets generalise better than others' and 'TONG and NEWS do not generalize as well' are not supported by the evidence shown. I suggest either adding repeated runs with confidence intervals, or explicitly describing the figure as a single illustrative run and softening the comparative claims.
minor comments (5)
  1. [Section 3] The text 'decide weather the marked expression' should be 'decide whether the marked expression is used metaphorically or literally.'
  2. [Section 3, 'Unified input format'] The sentence listing multilingual examples contains two unresolved placeholders: 'Mandarin Chinese sentences in ?' and 'Farsi sentences in ?'. Please cite the relevant datasets or remove these examples.
  3. [Appendix A] The JANK entry contains a duplicated word: 'because because they are not used.'
  4. [Table 4] The license for TSV_A is given as 'see data page' without a URL in the table; please include the actual license name or a stable link.
  5. [Section 4.4] The phrase 'restricted search for a label' is ambiguous; consider rephrasing to 'the label can be filtered to metaphorical, literal, anomalous, or other categories.'

Circularity Check

0 steps flagged · score 0.0 of 10

No circularity: the paper is a system/repository description, and the case-study evaluation and self-citations are illustrative or methodological, not load-bearing.

full rationale

MetaphorShare is a system/website paper; there is no derivation chain whose conclusion is equivalent to its input. The unified CSV format with tags <m>, <l>, <t>, <a>, <u> and free columns is presented as a flexible design choice, not as a result derived from the datasets themselves; the claim that information is preserved is an assumption, and the lack of round-trip verification is a completeness/correctness concern, not a circular step. The evaluation in Section 5 is explicitly a case study ('to illustrate a possible usage') and a sanity check of uploading, database insertion, and search; it does not claim to validate the repository or a scientific theory. VUAC_BO (Boisson et al., 2023) is an author-created dataset among 25 resources, and the hyperparameter configuration is reused from that prior work ('Similarly to the experiments in Boisson et al. (2023)'), but this is a methodological reference, not an argument whose conclusion depends on the prior paper's claims. No uniqueness theorem, ansatz, fitted parameter, or prediction is imported from self-citations. Even if the format-preservation assumption proves false, that would be an unsupported empirical claim, not a circular derivation. Therefore no circularity is exhibited, and the score is 0.

Assumptions & free parameters 0 free parameters · 3 assumptions · 0 invented entities

The central claim depends on the adequacy of the proposed unified annotation format and on the legal right to redistribute third-party datasets. No numerical constants are fitted to support the platform's design; the experimental free choices in Section 5 (e.g., 800 training examples per dataset, BOHB hyperparameters) affect only the illustrative case study, not the repository itself.

assumptions (3)
  • domain assumption The tag set <m>, <l>, <t>, <a>, <u> is sufficient to represent the annotation information of diverse metaphor datasets without meaningful loss.
    Invoked in Section 3 (Unified input format) where all datasets are converted to CSV with these tags; if the tag set is insufficient, the unified format would misrepresent datasets and the repository's core value would be undermined.
  • domain assumption Redistribution of the hosted datasets is legally permitted by their licenses or by author permission.
    The repository depends on redistribution rights. Table 4 lists PVC with 'no license' and Cardillo datasets with CC BY-NC, so the paper relies on permissions beyond the stated licenses to host and redistribute these resources.
  • domain assumption The source datasets' labels are reliable enough for unified reuse in cross-dataset experiments.
    Section 5's cross-dataset analysis inherits any inconsistencies, annotation errors, or differing annotation guidelines present in the original datasets, which the repository does not correct.

how reviews work

0 comments
Cite this review

Pith. "Pith review of MetaphorShare: A Dynamic Collaborative Repository of Open Metaphor Datasets." pith.science (2026). https://pith.science/paper/NSBU6P63

@misc{pith2026241118260,
  author       = {Pith},
  title        = {Pith review of: MetaphorShare: A Dynamic Collaborative Repository of Open Metaphor Datasets},
  year         = {2026},
  howpublished = {\url{https://pith.science/paper/NSBU6P63}},
  note         = {Machine review of arXiv:2411.18260}
}
read the original abstract

The metaphor studies community has developed numerous valuable labelled corpora in various languages over the years. Many of these resources are not only unknown to the NLP community, but are also often not easily shared among the researchers. Both in human sciences and in NLP, researchers could benefit from a centralised database of labelled resources, easily accessible and unified under an identical format. To facilitate this, we present MetaphorShare, a website to integrate metaphor datasets making them open and accessible. With this effort, our aim is to encourage researchers to share and upload more datasets in any language in order to facilitate metaphor studies and the development of future metaphor processing NLP systems. The website has four main functionalities: upload, download, search and label metaphor datasets. It is accessible at www.metaphorshare.com.

Figures

Figures reproduced from arXiv: 2411.18260 by the authors.

Figure 1
Figure 1. METAPHORSHARE search page. Specific datasets, languages, and tag types can be selected, and a text-based search within tagged expressions or into the entire text is implemented. Additional features provided with the record appear when clicking the Show Details button. tions between AI/NLP communities and linguis￾tics/metaphor studies by facilitating the unification of dataset formats and the access to existing re￾so… view at source ↗
Figure 2
Figure 2. Screenshot of the online annotation tool showing the text input area, tag selection and creation, and [PITH_FULL_IMAGE:figures/full_fig_p004_2.png] view at source ↗
Figure 3
Figure 3. Results of the cross dataset evaluation. F1-score of the [PITH_FULL_IMAGE:figures/full_fig_p007_3.png] view at source ↗
Figures from the paper (4 more)
Figure 4
Figure 4. Figure 4: Screenshot of top of the the datasets information page in the catalog section of the website. The English [PITH_FULL_IMAGE:figures/full_fig_p012_4.png]
Figure 5
Figure 5. Figure 5: Screenshot of the file format check for a rejected file. The line the error occurs in the CSV file and the [PITH_FULL_IMAGE:figures/full_fig_p013_5.png]
Figure 6
Figure 6. Figure 6: Screenshot of the file format check after a CSV file is accepted for manual review. [PITH_FULL_IMAGE:figures/full_fig_p013_6.png]
Figure 7
Figure 7. Figure 7: Screenshot showing dataset rows available for tagging or edition, as displayed in the annotation tool. [PITH_FULL_IMAGE:figures/full_fig_p013_7.png]

Discussion (0). Continue with ORCID to comment.

Reference graph

Works this paper leans on

66 extracted references · 34 canonical work pages

  1. [1]

    online" 'onlinestring :=

    ENTRY address archivePrefix author booktitle chapter edition editor eid eprint eprinttype howpublished institution journal key month note number organization pages publisher school series title type volume year doi pubmed url lastchecked label extra.label sort.label short.list INTEGERS output.state before.all mid.sentence after.sentence after.block STRING...

  2. [2]

    write newline

    " write newline "" before.all 'output.state := FUNCTION n.dashify 't := "" t empty not t #1 #1 substring "-" = t #1 #2 substring "--" = not "--" * t #2 global.max substring 't := t #1 #1 substring "-" = "-" * t #2 global.max substring 't := while if t #1 #1 substring * t #2 global.max substring 't := if while FUNCTION word.in bbl.in capitalize " " * FUNCT...

  3. [3]

    Lasha Abzianidze, Johannes Bjerva, Kilian Evang, Hessel Haagsma, Rik van Noord, Pierre Ludmann, Duc-Duy Nguyen, and Johan Bos. 2017. https://aclanthology.org/E17-2039 The P arallel M eaning B ank: Towards a multilingual corpus of translations annotated with compositional meaning representations . In Proceedings of the 15th Conference of the E uropean Chap...

  4. [4]

    Rodrigo Agerri, John Barnden, Mark Lee, and Alan Wallington. 2008. https://aclanthology.org/W08-2228 Textual entailment as an evaluation framework for metaphor resolution: A proposal . In Semantics in Text Processing. STEP 2008 Conference Proceedings , pages 357--363. College Publications

  5. [5]

    Daniel Baleato Rodr \' guez, Verna Dankers, Preslav Nakov, and Ekaterina Shutova. 2023. https://doi.org/10.18653/v1/2023.findings-eacl.35 Paper bullets: Modeling propaganda with the help of metaphor . In Findings of the Association for Computational Linguistics: EACL 2023, pages 472--489, Dubrovnik, Croatia. Association for Computational Linguistics

  6. [6]

    Julia Birke and Anoop Sarkar. 2006. http://aclweb.org/anthology/E06-1042 A clustering approach for nearly unsupervised recognition of nonliteral language . In 11th Conference of the European Chapter of the Association for Computational Linguistics

  7. [7]

    Joanne Boisson, Luis Espinosa-Anke, and Jose Camacho-Collados. 2023. https://doi.org/10.18653/v1/2023.emnlp-main.406 Construction artifacts in metaphor identification datasets . In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, pages 6581--6590, Singapore. Association for Computational Linguistics

  8. [8]

    Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert - Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litw...

Show all 66 references
  1. [9]

    Cameron and R

    L. Cameron and R. Maslen. 2010. https://books.google.fr/books?id=3nL9PgAACAAJ Metaphor Analysis: Research Practice in Applied Linguistics, Social Sciences and the Humanities . Studies in applied linguistics. Equinox

  2. [10]

    Cardillo, Christine Watson, and Anjan Chatterjee

    Eileen R. Cardillo, Christine Watson, and Anjan Chatterjee. 2017. https://doi.org/10.3758/s13428-016-0717-1 Stimulus needs are a moving target: 240 additional matched literal and metaphorical sentences for testing neural hypotheses about metaphor . Behavior Research Methods, 4...

  3. [11]

    Cardillo, G.L

    E.R. Cardillo, G.L. Schmidt, A. Kranjec, and A. Chatterjee. 2010. https://doi.org/10.3758/BRM.42.3.651 Stimulus design is an obstacle course: 560 matched literal and metaphorical sentences for testing neural hypotheses about metaphor

  4. [12]

    Tuhin Chakrabarty, Debanjan Ghosh, Adam Poliak, and Smaranda Muresan. 2021 a . https://doi.org/10.18653/v1/2021.findings-acl.297 Figurative language in recognizing textual entailment . In Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021, pages 3354--3...

  5. [13]

    Tuhin Chakrabarty, Xurui Zhang, Smaranda Muresan, and Nanyun Peng. 2021 b . https://doi.org/10.18653/v1/2021.naacl-main.336 MERMAID : Metaphor generation with symbolism and discriminative decoding . In Proceedings of the 2021 Conference of the North American Chapter of the Ass...

  6. [14]

    Jonathan Dunn. 2014. https://doi.org/10.3115/v1/P14-2121 Measuring metaphoricity . In Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), pages 745--751, Baltimore, Maryland. Association for Computational Linguistics

  7. [15]

    Stefan Falkner, Aaron Klein, and Frank Hutter. 2018. https://arxiv.org/abs/1807.01774 Bohb: Robust and efficient hyperparameter optimization at scale . Preprint, arXiv:1807.01774

  8. [16]

    Dan Fass. 1997. Processing metaphor and metonymy

  9. [17]

    Adriano Ferraresi, Eros Zanchetta, Marco Baroni, and Silvia Bernardini. 2008. https://api.semanticscholar.org/CorpusID:4847295 Introducing and evaluating ukwac , a very large web-derived corpus of english

  10. [18]

    Mengshi Ge, Rui Mao, and Erik Cambria. 2023. https://doi.org/10.1007/s10462-023-10564-7 A survey on computational metaphor processing techniques: From identification, interpretation, generation to application . Artificial Intelligence Review, 56(02):1829--1895

  11. [19]

    Debanjan Ghosh, Beata Beigman Klebanov, Smaranda Muresan, Anna Feldman, Soujanya Poria, and Tuhin Chakrabarty, editors. 2022. https://aclanthology.org/2022.flp-1.0 Proceedings of the 3rd Workshop on Figurative Language Processing (FLP) . Association for Computational Linguisti...

  12. [20]

    Debanjan Ghosh, Smaranda Muresan, Anna Feldman, Tuhin Chakrabarty, and Emmy Liu, editors. 2024. https://aclanthology.org/2024.figlang-1.0 Proceedings of the 4th Workshop on Figurative Language Processing (FigLang 2024) . Association for Computational Linguistics, Mexico City, ...

  13. [21]

    Jonathan Gordon, Jerry Hobbs, Jonathan May, Michael Mohler, Fabrizio Morbini, Bryan Rink, Marc Tomlinson, and Suzanne Wertheim. 2015. https://doi.org/10.3115/v1/W15-1407 A corpus of rich metaphor annotation . In Proceedings of the Third Workshop on Metaphor in NLP , pages 56--...

  14. [22]

    Dario Guti \'e rrez, Ekaterina Shutova, Tyler Marghetis, and Benjamin Bergen

    E. Dario Guti \'e rrez, Ekaterina Shutova, Tyler Marghetis, and Benjamin Bergen. 2016. https://doi.org/10.18653/v1/P16-1018 Literal and metaphorical senses in compositional distributional semantic models . In Proceedings of the 54th Annual Meeting of the Association for Comput...

  15. [23]

    Hessel Haagsma, Johan Bos, and Malvina Nissim. 2020. https://aclanthology.org/2020.lrec-1.35 MAGPIE : A large corpus of potentially idiomatic expressions . In Proceedings of the Twelfth Language Resources and Evaluation Conference, pages 279--287, Marseille, France. European L...

  16. [24]

    Sooji Han, Rui Mao, and Erik Cambria. 2022. https://aclanthology.org/2022.coling-1.9 Hierarchical attention network for explainable depression detection on T witter aided by metaphor concept mappings . In Proceedings of the 29th International Conference on Computational Lingui...

  17. [25]

    Katarzyna Jankowiak. 2020. https://doi.org/10.1007/s10936-020-09695-7 Normative data for novel nominal metaphors, novel similes, literal, and anomalous utterances in polish and english . Journal of Psycholinguistic Research, 49(4):541--569

  18. [26]

    Rohan Joseph, Timothy Liu, Aik Beng Ng, Simon See, and Sunny Rai. 2023. https://doi.org/10.18653/v1/2023.findings-acl.641 N ews M et : A ` do it all ' dataset of contemporary metaphors in news headlines . In Findings of the Association for Computational Linguistics: ACL 2023, ...

  19. [27]

    Nina Julich-Warpakowski. 2022. https://doi.org/10.1075/milcc.10 Motion Metaphors in Music Criticism: An empirical investigation of their conceptual motivation and their metaphoricity . Metaphor in Language, Cognition, and Communication. John Benjamins Publishing Company

  20. [28]

    Albert Katz, Allan Paivio, Marc Marschark, and Jim Clark. 1988. https://doi.org/10.1207/s15327868ms0304_1 Norms for 204 literary and 260 nonliterary metaphors on 10 psychological dimensions . Metaphor and Symbol - METAPHOR SYMB, 3:191--214

  21. [29]

    Adam Kilgarriff, Vít Baisa, Jan Bušta, Miloš Jakubíček, Vojtěch Kovář, Jan Michelfeit, Pavel Rychlý, and Vít Suchomel. 2014. The sketch engine: ten years on. Lexicography, 1:7--36

  22. [30]

    Walter Kintsch. 2000. https://doi.org/10.3758/BF03212981 Metaphor comprehension: A computational theory . Psychonomic Bulletin & Review , 7(2):257--266

  23. [31]

    Veronika Koller, Andrew Hardie, Paul Rayson, Elena Semino, and Lancaster. 2008. Using a semantic annotation tool for the analysis of metaphor in discourse. Metaphorik.De, 15

  24. [32]

    Shreyas Kulkarni, Arkadiy Saakyan, Tuhin Chakrabarty, and Smaranda Muresan. 2024. https://aclanthology.org/2024.figlang-1.16 A report on the F ig L ang 2024 shared task on multimodal figurative language . In Proceedings of the 4th Workshop on Figurative Language Processing (Fi...

  25. [33]

    G. Lakoff. 1994. https://books.google.fr/books?id=lGSyPgAACAAJ Master Metaphor List . University of California

  26. [34]

    George Lakoff and Mark Johnson. 1980. Metaphors we Live by. University of Chicago Press, Chicago

  27. [35]

    Chee Wee (Ben) Leong, Beata Beigman Klebanov, and Ekaterina Shutova. 2018. https://doi.org/10.18653/v1/W18-0907 A report on the 2018 VUA metaphor detection shared task . In Proceedings of the Workshop on Figurative Language Processing, pages 56--66, New Orleans, Louisiana. Ass...

  28. [36]

    Yucheng Li, Frank Guerin, and Chenghua Lin. 2024. https://arxiv.org/abs/2401.16012 Finding challenging metaphors that confuse pretrained language models . Preprint, arXiv:2401.16012

  29. [37]

    Gonzalez, and Ion Stoica

    Richard Liaw, Eric Liang, Robert Nishihara, Philipp Moritz, Joseph E. Gonzalez, and Ion Stoica. 2018. https://arxiv.org/abs/1807.05118 Tune: A research platform for distributed model selection and training . Preprint, arXiv:1807.05118

  30. [38]

    Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019. Roberta: A robustly optimized bert pretraining approach. arXiv preprint arXiv:1907.11692

  31. [39]

    Rui Mao, Xiao Li, Kai He, Mengshi Ge, and Erik Cambria. 2023. https://doi.org/10.18653/v1/2023.acl-demo.12 M eta P ro online: A computational metaphor processing online system . In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume ...

  32. [40]

    Rui Mao, Chenghua Lin, and Frank Guerin. 2018. https://doi.org/10.18653/v1/P18-1113 Word embedding and W ord N et based metaphor identification and interpretation . In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Paper...

  33. [41]

    Rui Mao, Chenghua Lin, and Frank Guerin. 2019. https://doi.org/10.18653/v1/P19-1378 End-to-end sequential metaphor identification inspired by linguistic theories . In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, pages 3888--3898, Flo...

  34. [42]

    James H. Martin. 1990. A Computational Model of Metaphor Interpretation. Academic Press Professional, Inc., San Diego, CA, USA

  35. [43]

    Saif Mohammad, Ekaterina Shutova, and Peter Turney. 2016. https://doi.org/10.18653/v1/S16-2003 Metaphor as a medium for emotion: An empirical study . In Proceedings of the Fifth Joint Conference on Lexical and Computational Semantics, pages 23--33, Berlin, Germany. Association...

  36. [44]

    Michael Mohler, Mary Brunson, Bryan Rink, and Marc Tomlinson. 2016. https://www.aclweb.org/anthology/L16-1668 Introducing the LCC metaphor datasets . In Proceedings of the Tenth International Conference on Language Resources and Evaluation ( LREC '16) , pages 4221--4227, Porto...

  37. [45]

    Susan Nacey. 2022. https://doi.org/10.18710/95QQ2W Replication data for: Systematic metaphors in Norwegian doctoral dissertation acknowledgements

  38. [46]

    Dorst, Tina Krennmayr, and W

    Susan Nacey, Aletta G. Dorst, Tina Krennmayr, and W. Gudrun Reijnierse, editors. 2019 a . https://www.jbe-platform.com/content/books/9789027261755 Metaphor Identification in Multiple Languages: MIPVU around the world . John Benjamins

  39. [47]

    Dorst, and W

    Susan Nacey, Tina Krennmayr, Aletta G. Dorst, and W. Gudrun Reijnierse. 2019 b . https://doi.org/10.18710/F04UW5 Replication Data for: What the MIPVU protocol doesn’t tell you (even though it really does)

  40. [48]

    Hiroki Nakayama, Takahiro Kubo, Junya Kamura, Yasufumi Taniguchi, and Xu Liang. 2018. https://github.com/doccano/doccano doccano : Text annotation tool for human . Software available from https://github.com/doccano/doccano

  41. [49]

    Natalie Parde and Rodney Nielsen. 2018. https://doi.org/10.1609/aaai.v32i1.11940 Exploring the terrain of metaphor novelty: A regression-based approach for automatically scoring metaphors . Proceedings of the AAAI Conference on Artificial Intelligence, 32(1)

  42. [50]

    Jiaxin Pei, Aparna Ananthasubramaniam, Xingyao Wang, Naitian Zhou, Apostolos Dedeloudis, Jackson Sargent, and David Jurgens. 2022. https://doi.org/10.18653/v1/2022.emnlp-demos.33 POTATO : The portable text annotation tool . In Proceedings of the 2022 Conference on Empirical Me...

  43. [51]

    Richards

    I.A. Richards. 1936. https://books.google.fr/books?id=AEe-snyOyU4C The Philosophy of Rhetoric . Bryn Mawr College. Mary Flexner lectures. Oxford University Press

  44. [52]

    o rn Gamb \

    Chhavi Sharma, Deepesh Bhageria, William Scott, Srinivas PYKL, Amitava Das, Tanmoy Chakraborty, Viswanath Pulabaigari, and Bj \"o rn Gamb \"a ck. 2020. https://doi.org/10.18653/v1/2020.semeval-1.99 S em E val-2020 task 8: Memotion analysis- the visuo-lingual metaphor! In Proce...

  45. [53]

    Gerard Steen. 2010. A method for linguistic metaphor identification: from MIP to MIPVU, volume v. 14 of Converging evidence in language and communication research. John Benjamins Pub. Co., Amsterdam

  46. [54]

    Kevin Stowe, Tuhin Chakrabarty, Nanyun Peng, Smaranda Muresan, and Iryna Gurevych. 2021. https://doi.org/10.18653/v1/2021.acl-long.524 Metaphor generation with conceptual mappings . In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and ...

  47. [55]

    Kevin Stowe, Prasetya Utama, and Iryna Gurevych. 2022. https://doi.org/10.18653/v1/2022.acl-long.369 IMPLI : Investigating NLI models ' performance on figurative language . In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Lo...

  48. [56]

    Harish Tayyar Madabushi, Edward Gow-Smith, Marcos Garcia, Carolina Scarton, Marco Idiart, and Aline Villavicencio. 2022 a . https://doi.org/10.18653/v1/2022.semeval-1.13 S em E val-2022 task 2: Multilingual idiomaticity detection and sentence embedding . In Proceedings of the ...

  49. [57]

    Harish Tayyar Madabushi, Edward Gow-Smith, Marcos Garcia, Carolina Scarton, Marco Idiart, and Aline Villavicencio. 2022 b . SemEval-2022 Task 2 : Multilingual Idiomaticity Detection and Sentence Embedding . In Proceedings of the 16th International Workshop on Semantic Evaluati...

  50. [58]

    Xiaoyu Tong, Rochelle Choenni, Martha Lewis, and Ekaterina Shutova. 2024. https://arxiv.org/abs/2403.11810 Metaphor understanding challenge dataset for llms . Preprint, arXiv:2403.11810

  51. [59]

    Yulia Tsvetkov, Leonid Boytsov, Anatole Gershman, Eric Nyberg, and Chris Dyer. 2014. https://doi.org/10.3115/v1/P14-1024 Metaphor detection with cross-lingual model transfer . In Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (Volume 1:...

  52. [60]

    Yuancheng Tu and Dan Roth. 2012. Sorting out the most confusing english phrasal verbs. In Proceedings of the First Joint Conference on Lexical and Computational Semantics

  53. [61]

    Peter Turney, Yair Neuman, Dan Assaf, and Yohai Cohen. 2011. https://aclanthology.org/D11-1063 Literal and metaphorical sense identification through concrete and abstract context . In Proceedings of the 2011 Conference on Empirical Methods in Natural Language Processing, pages...

  54. [62]

    Tony Veale. 2016. https://doi.org/10.18653/v1/W16-1105 Round up the usual suspects: Knowledge-based metaphor generation . In Proceedings of the Fourth Workshop on Metaphor in NLP , pages 34--41, San Diego, California. Association for Computational Linguistics

  55. [63]

    Tony Veale and Guofu Li. 2012. https://aclanthology.org/P12-3002 Specifying viewpoint and information need with affective metaphors: A system demonstration of the metaphor-magnet web app/service . In Proceedings of the ACL 2012 System Demonstrations , pages 7--12, Jeju Island,...

  56. [64]

    Lennart Wachowiak and Dagmar Gromann. 2023. https://doi.org/10.18653/v1/2023.acl-long.58 Does GPT -3 grasp metaphors? identifying metaphor mappings with generative language models . In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Vol...

  57. [65]

    Yorick Wilks. 1973. https://apps.dtic.mil/sti/citations/AD0764652 Preference semantics

  58. [66]

    Ziheng Zeng and Suma Bhat. 2021. https://doi.org/10.1162/tacl_a_00442 Idiomatic expression identification using semantic compatibility . Transactions of the Association for Computational Linguistics, 9:1546--1562

Pith tools

Reviewed August 12, 2026 · model on record in the stance chip above.