Skip to main content

Showing 1–3 of 3 results for author: Mashetty, S

Searching in archive cs. Search in all archives.
.
  1. arXiv:2404.15522  [pdf, other

    cs.CL cs.AI

    LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

    Authors: Mihir Parmar, Nisarg Patel, Neeraj Varshney, Mutsumi Nakamura, Man Luo, Santosh Mashetty, Arindam Mitra, Chitta Baral

    Abstract: Recently developed large language models (LLMs) have been shown to perform remarkably well on a wide range of language understanding tasks. But, can they really "reason" over the natural language? This question has been receiving significant research attention and many reasoning skills such as commonsense, numerical, and qualitative have been studied. However, the crucial skill pertaining to 'logi… ▽ More

    Submitted 6 June, 2024; v1 submitted 23 April, 2024; originally announced April 2024.

    Comments: Accepted at ACL(Main) 2024 | First version available @ https://openreview.net/forum?id=7NR2ZVzZxx

  2. arXiv:2306.05539  [pdf, other

    cs.CL

    Instruction Tuned Models are Quick Learners

    Authors: Himanshu Gupta, Saurabh Arjun Sawant, Swaroop Mishra, Mutsumi Nakamura, Arindam Mitra, Santosh Mashetty, Chitta Baral

    Abstract: Instruction tuning of language models has demonstrated the ability to enhance model generalization to unseen tasks via in-context learning using a few examples. However, typical supervised learning still requires a plethora of downstream training data for finetuning. Often in real-world situations, there is a scarcity of data available for finetuning, falling somewhere between few shot inference a… ▽ More

    Submitted 17 May, 2023; originally announced June 2023.

    Comments: 9 pages, 5 figures, 19 Tables (inclusing appendix), 12 pages of Appendix

  3. arXiv:2109.08079  [pdf, other

    cs.IR cs.CL cs.LG

    Context-NER : Contextual Phrase Generation at Scale

    Authors: Himanshu Gupta, Shreyas Verma, Santosh Mashetty, Swaroop Mishra

    Abstract: Named Entity Recognition (NER) has seen significant progress in recent years, with numerous state-of-the-art (SOTA) models achieving high performance. However, very few studies have focused on the generation of entities' context. In this paper, we introduce CONTEXT-NER, a task that aims to generate the relevant context for entities in a sentence, where the context is a phrase describing the entity… ▽ More

    Submitted 8 June, 2023; v1 submitted 16 September, 2021; originally announced September 2021.

    Comments: 29 pages, 5 Figures, 2 AlgorithmS, 17 Tables. Accepted in NeurIPS 2022 - Efficient Natural Language and Speech Processing (ENLSP) Workshop