Skip to main content

Showing 1–7 of 7 results for author: Hahn-Powell, G

Searching in archive cs. Search in all archives.
.
  1. arXiv:2208.06957  [pdf, other

    cs.CL cs.AI cs.LG

    Syntax-driven Data Augmentation for Named Entity Recognition

    Authors: Arie Pratama Sutiono, Gus Hahn-Powell

    Abstract: In low resource settings, data augmentation strategies are commonly leveraged to improve performance. Numerous approaches have attempted document-level augmentation (e.g., text classification), but few studies have explored token-level augmentation. Performed naively, data augmentation can produce semantically incongruent and ungrammatical examples. In this work, we compare simple masked language… ▽ More

    Submitted 1 October, 2022; v1 submitted 14 August, 2022; originally announced August 2022.

    Comments: submitted to Pattern-based Approaches to NLP in the Age of Deep Learning 2022 (Pan-DL 2022)

    MSC Class: 68T50 ACM Class: I.2.7

  2. arXiv:2107.05133  [pdf, ps, other

    cs.CL

    Computer-assisted construct classification of organizational performance concerning different stakeholder groups

    Authors: Seethalakshmi Gopalakrishnan, Victor Chen, Gus Hahn-Powell, Bharadwaj Tirunagar

    Abstract: The number of research articles in business and management has dramatically increased along with terminology, constructs, and measures. Proper classification of organizational performance constructs from research articles plays an important role in categorizing the literature and understanding to whom its research implications may be relevant. In this work, we classify constructs (i.e., concepts a… ▽ More

    Submitted 23 August, 2021; v1 submitted 11 July, 2021; originally announced July 2021.

  3. arXiv:1711.00529  [pdf, other

    cs.CL

    Text Annotation Graphs: Annotating Complex Natural Language Phenomena

    Authors: Angus G. Forbes, Kristine Lee, Gus Hahn-Powell, Marco A. Valenzuela-Escárcega, Mihai Surdeanu

    Abstract: This paper introduces a new web-based software tool for annotating text, Text Annotation Graphs, or TAG. It provides functionality for representing complex relationships between words and word phrases that are not available in other software tools, including the ability to define and visualize relationships between the relationships themselves (semantic hypergraphs). Additionally, we include an ap… ▽ More

    Submitted 1 March, 2018; v1 submitted 1 November, 2017; originally announced November 2017.

    Comments: Accepted to LREC'18, http://lrec2018.lrec-conf.org/en/conference-programme/accepted-papers/

  4. arXiv:1606.09604  [pdf, other

    cs.CL

    SnapToGrid: From Statistical to Interpretable Models for Biomedical Information Extraction

    Authors: Marco A. Valenzuela-Escarcega, Gus Hahn-Powell, Dane Bell, Mihai Surdeanu

    Abstract: We propose an approach for biomedical information extraction that marries the advantages of machine learning models, e.g., learning directly from data, with the benefits of rule-based approaches, e.g., interpretability. Our approach starts by training a feature-based statistical model, then converts this model to a rule-based variant by converting its features to rules, and "snapping to grid" the… ▽ More

    Submitted 30 June, 2016; originally announced June 2016.

  5. arXiv:1606.08089  [pdf, other

    cs.CL

    This before That: Causal Precedence in the Biomedical Domain

    Authors: Gus Hahn-Powell, Dane Bell, Marco A. Valenzuela-Escárcega, Mihai Surdeanu

    Abstract: Causal precedence between biochemical interactions is crucial in the biomedical domain, because it transforms collections of individual interactions, e.g., bindings and phosphorylations, into the causal mechanisms needed to inform meaningful search and inference. Here, we analyze causal precedence in the biomedical domain as distinct from open-domain, temporal precedence. First, we describe a nove… ▽ More

    Submitted 26 June, 2016; originally announced June 2016.

    Comments: To appear in the proceedings of the 2016 Workshop on Biomedical Natural Language Processing (BioNLP 2016)

  6. arXiv:1603.03758  [pdf, other

    cs.CL

    Sieve-based Coreference Resolution in the Biomedical Domain

    Authors: Dane Bell, Gus Hahn-Powell, Marco A. Valenzuela-Escárcega, Mihai Surdeanu

    Abstract: We describe challenges and advantages unique to coreference resolution in the biomedical domain, and a sieve-based architecture that leverages domain knowledge for both entity and event coreference resolution. Domain-general coreference resolution algorithms perform poorly on biomedical documents, because the cues they rely on such as gender are largely absent in this domain, and because they do n… ▽ More

    Submitted 2 September, 2016; v1 submitted 11 March, 2016; originally announced March 2016.

    Comments: This paper appears in LREC 2016

  7. arXiv:1509.07513  [pdf, other

    cs.CL

    Description of the Odin Event Extraction Framework and Rule Language

    Authors: Marco A. Valenzuela-Escárcega, Gus Hahn-Powell, Mihai Surdeanu

    Abstract: This document describes the Odin framework, which is a domain-independent platform for developing rule-based event extraction models. Odin aims to be powerful (the rule language allows the modeling of complex syntactic structures) and robust (to recover from syntactic parsing errors, syntactic patterns can be freely mixed with surface, token-based patterns), while remaining simple (some domain gra… ▽ More

    Submitted 24 September, 2015; originally announced September 2015.