ACLM: A Selective-Denoising based Generative Data Augmentation Approach for Low-Resource Complex NER
AI-generated Key Points
- Complex Named Entity Recognition (NER) is a challenging task in low-context text
- Existing data augmentation techniques for low-resource complex NER suffer from context-entity mismatch
- ACLM (Attention-map aware keyword selection for Conditional Language Model fine-tuning) is a novel data augmentation approach based on conditional generation
- ACLM uses selective masking aided by attention maps to retain named entities and relevant keywords
- ACLM outperforms other neural baselines by a significant margin in monolingual, cross-lingual, and multilingual complex NER across various low-resource settings
- ACLM also demonstrates promising results in biomedical research and other domains suffering from data scarcity
- Mixner algorithm enhances diversity of augmentations by mixing two templates during generation phase
- Benefits of ACLM demonstrated qualitatively and quantitatively on the Multi-CoNER dataset
- ACLM effectively addresses context-entity mismatch problem and generates diverse, coherent, and high-quality augmentations
- ACLM outperforms previous methods in science and medicine domains with absolute gains ranging from 1% to 11%
Authors: Sreyan Ghosh, Utkarsh Tyagi, Manan Suri, Sonal Kumar, S Ramaneswaran, Dinesh Manocha
Abstract: Complex Named Entity Recognition (NER) is the task of detecting linguistically complex named entities in low-context text. In this paper, we present ACLM Attention-map aware keyword selection for Conditional Language Model fine-tuning), a novel data augmentation approach based on conditional generation to address the data scarcity problem in low-resource complex NER. ACLM alleviates the context-entity mismatch issue, a problem existing NER data augmentation techniques suffer from and often generates incoherent augmentations by placing complex named entities in the wrong context. ACLM builds on BART and is optimized on a novel text reconstruction or denoising task - we use selective masking (aided by attention maps) to retain the named entities and certain keywords in the input sentence that provide contextually relevant additional knowledge or hints about the named entities. Compared with other data augmentation strategies, ACLM can generate more diverse and coherent augmentations preserving the true word sense of complex entities in the sentence. We demonstrate the effectiveness of ACLM both qualitatively and quantitatively on monolingual, cross-lingual, and multilingual complex NER across various low-resource settings. ACLM outperforms all our neural baselines by a significant margin (1%-36%). In addition, we demonstrate the application of ACLM to other domains that suffer from data scarcity (e.g., biomedical). In practice, ACLM generates more effective and factual augmentations for these domains than prior methods. Code: https://github.com/Sreyan88/ACLM
Ask questions about this paper to our AI assistant
You can also chat with multiple papers at once here.
Assess the quality of the AI-generated content by voting
Score: 0
Why do we need votes?
Votes are used to determine whether we need to re-run our summarizing tools. If the count reaches -10, our tools can be restarted.
The previous summary was created more than a year ago and can be re-run (if necessary) by clicking on the Run button below.
Similar papers summarized with our AI tools
Navigate through even more similar papers through a
tree representationLook for similar papers (in beta version)
By clicking on the button above, our algorithm will scan all papers in our database to find the closest based on the contents of the full papers and not just on metadata. Please note that it only works for papers that we have generated summaries for and you can rerun it from time to time to get a more accurate result while our database grows.
Disclaimer: The AI-based summarization tool and virtual assistant provided on this website may not always provide accurate and complete summaries or responses. We encourage you to carefully review and evaluate the generated content to ensure its quality and relevance to your needs.