Skip to main content

Showing 1–31 of 31 results for author: Sanyal, D

Searching in archive cs. Search in all archives.
.
  1. arXiv:2505.23788  [pdf, ps, other

    cs.CL cs.AI

    Nine Ways to Break Copyright Law and Why Our LLM Won't: A Fair Use Aligned Generation Framework

    Authors: Aakash Sen Sharma, Debdeep Sanyal, Priyansh Srivastava, Sundar Atreya H., Shirish Karande, Mohan Kankanhalli, Murari Mandal

    Abstract: Large language models (LLMs) commonly risk copyright infringement by reproducing protected content verbatim or with insufficient transformative modifications, posing significant ethical, legal, and practical concerns. Current inference-time safeguards predominantly rely on restrictive refusal-based filters, often compromising the practical utility of these models. To address this, we collaborated… ▽ More

    Submitted 25 May, 2025; originally announced May 2025.

    Comments: 30 Pages

  2. arXiv:2505.19173  [pdf, other

    cs.AI

    Investigating Pedagogical Teacher and Student LLM Agents: Genetic Adaptation Meets Retrieval Augmented Generation Across Learning Style

    Authors: Debdeep Sanyal, Agniva Maiti, Umakanta Maharana, Dhruv Kumar, Ankur Mali, C. Lee Giles, Murari Mandal

    Abstract: Effective teaching requires adapting instructional strategies to accommodate the diverse cognitive and behavioral profiles of students, a persistent challenge in education and teacher training. While Large Language Models (LLMs) offer promise as tools to simulate such complex pedagogical environments, current simulation frameworks are limited in two key respects: (1) they often reduce students to… ▽ More

    Submitted 25 May, 2025; originally announced May 2025.

    Comments: 38 Pages

  3. arXiv:2505.19165  [pdf, ps, other

    cs.AI

    OrgAccess: A Benchmark for Role Based Access Control in Organization Scale LLMs

    Authors: Debdeep Sanyal, Umakanta Maharana, Yash Sinha, Hong Ming Tan, Shirish Karande, Mohan Kankanhalli, Murari Mandal

    Abstract: Role-based access control (RBAC) and hierarchical structures are foundational to how information flows and decisions are made within virtually all organizations. As the potential of Large Language Models (LLMs) to serve as unified knowledge repositories and intelligent assistants in enterprise settings becomes increasingly apparent, a critical, yet under explored, challenge emerges: \textit{can th… ▽ More

    Submitted 17 June, 2025; v1 submitted 25 May, 2025; originally announced May 2025.

    Comments: 56 Pages

  4. arXiv:2503.18167  [pdf, other

    cs.CL cs.AI cs.LG

    Evaluating Negative Sampling Approaches for Neural Topic Models

    Authors: Suman Adhya, Avishek Lahiri, Debarshi Kumar Sanyal, Partha Pratim Das

    Abstract: Negative sampling has emerged as an effective technique that enables deep learning models to learn better representations by introducing the paradigm of learn-to-compare. The goal of this approach is to add robustness to deep learning models to learn better representation by comparing the positive samples against the negative ones. Despite its numerous demonstrations in various areas of computer v… ▽ More

    Submitted 25 March, 2025; v1 submitted 23 March, 2025; originally announced March 2025.

    Comments: Code is available at: https://github.com/AdhyaSuman/Eval_NegTM

    Journal ref: in IEEE Transactions on Artificial Intelligence, vol. 5, no. 11, pp. 5630-5642, Nov. 2024

  5. arXiv:2502.19339  [pdf, ps, other

    cs.CL

    Evaluating LLMs and Pre-trained Models for Text Summarization Across Diverse Datasets

    Authors: Tohida Rehman, Soumabha Ghosh, Kuntal Das, Souvik Bhattacharjee, Debarshi Kumar Sanyal, Samiran Chattopadhyay

    Abstract: Text summarization plays a crucial role in natural language processing by condensing large volumes of text into concise and coherent summaries. As digital content continues to grow rapidly and the demand for effective information retrieval increases, text summarization has become a focal point of research in recent years. This study offers a thorough evaluation of four leading pre-trained and open… ▽ More

    Submitted 13 March, 2025; v1 submitted 26 February, 2025; originally announced February 2025.

    Comments: 5 pages, 2 figures, 6 tables

  6. arXiv:2502.00406  [pdf, other

    cs.AI cs.CL

    ALU: Agentic LLM Unlearning

    Authors: Debdeep Sanyal, Murari Mandal

    Abstract: Information removal or suppression in large language models (LLMs) is a desired functionality, useful in AI regulation, legal compliance, safety, and privacy. LLM unlearning methods aim to remove information on demand from LLMs. Current LLM unlearning methods struggle to balance the unlearning efficacy and utility due to the competing nature of these objectives. Keeping the unlearning process comp… ▽ More

    Submitted 1 February, 2025; originally announced February 2025.

  7. arXiv:2501.15398  [pdf, other

    cs.CL

    How Green are Neural Language Models? Analyzing Energy Consumption in Text Summarization Fine-tuning

    Authors: Tohida Rehman, Debarshi Kumar Sanyal, Samiran Chattopadhyay

    Abstract: Artificial intelligence systems significantly impact the environment, particularly in natural language processing (NLP) tasks. These tasks often require extensive computational resources to train deep neural networks, including large-scale language models containing billions of parameters. This study analyzes the trade-offs between energy consumption and performance across three neural language mo… ▽ More

    Submitted 14 March, 2025; v1 submitted 25 January, 2025; originally announced January 2025.

  8. arXiv:2409.14602  [pdf, other

    cs.CL cs.AI

    Can pre-trained language models generate titles for research papers?

    Authors: Tohida Rehman, Debarshi Kumar Sanyal, Samiran Chattopadhyay

    Abstract: The title of a research paper communicates in a succinct style the main theme and, sometimes, the findings of the paper. Coming up with the right title is often an arduous task, and therefore, it would be beneficial to authors if title generation can be automated. In this paper, we fine-tune pre-trained language models to generate titles of papers from their abstracts. Additionally, we use GPT-3.5… ▽ More

    Submitted 13 October, 2024; v1 submitted 22 September, 2024; originally announced September 2024.

  9. Transfer Learning and Transformer Architecture for Financial Sentiment Analysis

    Authors: Tohida Rehman, Raghubir Bose, Samiran Chattopadhyay, Debarshi Kumar Sanyal

    Abstract: Financial sentiment analysis allows financial institutions like Banks and Insurance Companies to better manage the credit scoring of their customers in a better way. Financial domain uses specialized mechanisms which makes sentiment analysis difficult. In this paper, we propose a pre-trained language model which can help to solve this problem with fewer labelled data. We extend on the principles o… ▽ More

    Submitted 28 April, 2024; originally announced May 2024.

    Comments: 12 pages, 9 figures

    Journal ref: Proceedings of International Conference on Computational Intelligence, Data Science and Cloud Computing: IEM-ICDC 2021,pages 17--27

  10. GINopic: Topic Modeling with Graph Isomorphism Network

    Authors: Suman Adhya, Debarshi Kumar Sanyal

    Abstract: Topic modeling is a widely used approach for analyzing and exploring large document collections. Recent research efforts have incorporated pre-trained contextualized language models, such as BERT embeddings, into topic modeling. However, they often neglect the intrinsic informational value conveyed by mutual dependencies between words. In this study, we introduce GINopic, a topic modeling framewor… ▽ More

    Submitted 17 February, 2025; v1 submitted 2 April, 2024; originally announced April 2024.

    Comments: Accepted as a long paper for NAACL 2024 main conference

  11. Automatic Recognition of Learning Resource Category in a Digital Library

    Authors: Soumya Banerjee, Debarshi Kumar Sanyal, Samiran Chattopadhyay, Plaban Kumar Bhowmick, Partha Pratim Das

    Abstract: Digital libraries often face the challenge of processing a large volume of diverse document types. The manual collection and tagging of metadata can be a time-consuming and error-prone task. To address this, we aim to develop an automatic metadata extractor for digital libraries. In this work, we introduce the Heterogeneous Learning Resources (HLR) dataset designed for document image classificatio… ▽ More

    Submitted 28 November, 2023; originally announced January 2024.

    Comments: 2 pages, 3 figures, Published in JCDL 21

  12. arXiv:2309.16781  [pdf, other

    cs.CL cs.IR cs.LG

    Hallucination Reduction in Long Input Text Summarization

    Authors: Tohida Rehman, Ronit Mandal, Abhishek Agarwal, Debarshi Kumar Sanyal

    Abstract: Hallucination in text summarization refers to the phenomenon where the model generates information that is not supported by the input source document. Hallucination poses significant obstacles to the accuracy and reliability of the generated summaries. In this paper, we aim to reduce hallucinated outputs or hallucinations in summaries of long-form text documents. We have used the PubMed dataset, w… ▽ More

    Submitted 28 September, 2023; originally announced September 2023.

    Comments: 9 pages, 1 figure, 1 table

  13. arXiv:2309.10532  [pdf, other

    cs.AI

    A Cognitively-Inspired Neural Architecture for Visual Abstract Reasoning Using Contrastive Perceptual and Conceptual Processing

    Authors: Yuan Yang, Deepayan Sanyal, James Ainooson, Joel Michelson, Effat Farhana, Maithilee Kunda

    Abstract: We introduce a new neural architecture for solving visual abstract reasoning tasks inspired by human cognition, specifically by observations that human abstract reasoning often interleaves perceptual and conceptual processing as part of a flexible, iterative, and dynamic cognitive process. Inspired by this principle, our architecture models visual abstract reasoning as an iterative, self-contrasti… ▽ More

    Submitted 20 October, 2023; v1 submitted 19 September, 2023; originally announced September 2023.

  14. arXiv:2307.01292  [pdf, other

    cs.CR cs.AI cs.LG

    Pareto-Secure Machine Learning (PSML): Fingerprinting and Securing Inference Serving Systems

    Authors: Debopam Sanyal, Jui-Tse Hung, Manav Agrawal, Prahlad Jasti, Shahab Nikkhoo, Somesh Jha, Tianhao Wang, Sibin Mohan, Alexey Tumanov

    Abstract: Model-serving systems have become increasingly popular, especially in real-time web applications. In such systems, users send queries to the server and specify the desired performance metrics (e.g., desired accuracy, latency). The server maintains a set of models (model zoo) in the back-end and serves the queries based on the specified metrics. This paper examines the security, specifically robust… ▽ More

    Submitted 6 August, 2023; v1 submitted 3 July, 2023; originally announced July 2023.

    Comments: 17 pages, 9 figures, 6 tables

  15. arXiv:2305.19445  [pdf, other

    cs.CV cs.AI

    A Computational Account Of Self-Supervised Visual Learning From Egocentric Object Play

    Authors: Deepayan Sanyal, Joel Michelson, Yuan Yang, James Ainooson, Maithilee Kunda

    Abstract: Research in child development has shown that embodied experience handling physical objects contributes to many cognitive abilities, including visual learning. One characteristic of such experience is that the learner sees the same object from several different viewpoints. In this paper, we study how learning signals that equate different viewpoints -- e.g., assigning similar representations to dif… ▽ More

    Submitted 30 May, 2023; originally announced May 2023.

  16. arXiv:2304.12730  [pdf, other

    cs.CL

    CitePrompt: Using Prompts to Identify Citation Intent in Scientific Papers

    Authors: Avishek Lahiri, Debarshi Kumar Sanyal, Imon Mukherjee

    Abstract: Citations in scientific papers not only help us trace the intellectual lineage but also are a useful indicator of the scientific significance of the work. Citation intents prove beneficial as they specify the role of the citation in a given context. In this paper, we present CitePrompt, a framework which uses the hitherto unexplored approach of prompt-based learning for citation intent classificat… ▽ More

    Submitted 3 May, 2023; v1 submitted 25 April, 2023; originally announced April 2023.

    Comments: Selected for publication at ACM/IEEE JOINT CONFERENCE ON DIGITAL LIBRARIES 2023

  17. arXiv:2304.00235  [pdf, other

    cs.CL

    What Does the Indian Parliament Discuss? An Exploratory Analysis of the Question Hour in the Lok Sabha

    Authors: Suman Adhya, Debarshi Kumar Sanyal

    Abstract: The TCPD-IPD dataset is a collection of questions and answers discussed in the Lower House of the Parliament of India during the Question Hour between 1999 and 2019. Although it is difficult to analyze such a huge collection manually, modern text analysis tools can provide a powerful means to navigate it. In this paper, we perform an exploratory analysis of the dataset. In particular, we present i… ▽ More

    Submitted 1 April, 2023; originally announced April 2023.

    Comments: Accepted at the workshop PoliticalNLP co-located with the conference LREC 2022

  18. arXiv:2303.15973  [pdf, other

    cs.CL cs.LG

    Do Neural Topic Models Really Need Dropout? Analysis of the Effect of Dropout in Topic Modeling

    Authors: Suman Adhya, Avishek Lahiri, Debarshi Kumar Sanyal

    Abstract: Dropout is a widely used regularization trick to resolve the overfitting issue in large feedforward neural networks trained on a small dataset, which performs poorly on the held-out test subset. Although the effectiveness of this regularization trick has been extensively studied for convolutional neural networks, there is a lack of analysis of it for unsupervised models and in particular, VAE-base… ▽ More

    Submitted 28 March, 2023; originally announced March 2023.

    Comments: Accepted at EACL 2023

  19. Improving Neural Topic Models with Wasserstein Knowledge Distillation

    Authors: Suman Adhya, Debarshi Kumar Sanyal

    Abstract: Topic modeling is a dominant method for exploring document collections on the web and in digital libraries. Recent approaches to topic modeling use pretrained contextualized language models and variational autoencoders. However, large neural topic models have a considerable memory footprint. In this paper, we propose a knowledge distillation framework to compress a contextualized topic model witho… ▽ More

    Submitted 20 June, 2024; v1 submitted 27 March, 2023; originally announced March 2023.

    Comments: Accepted at ECIR 2023

  20. arXiv:2303.14951  [pdf, other

    cs.CL cs.LG

    Improving Contextualized Topic Models with Negative Sampling

    Authors: Suman Adhya, Avishek Lahiri, Debarshi Kumar Sanyal, Partha Pratim Das

    Abstract: Topic modeling has emerged as a dominant method for exploring large document collections. Recent approaches to topic modeling use large contextualized language models and variational autoencoders. In this paper, we propose a negative sampling mechanism for a contextualized topic model to improve the quality of the generated topics. In particular, during model training, we perturb the generated doc… ▽ More

    Submitted 27 March, 2023; originally announced March 2023.

    Comments: Accepted at 19th International Conference on Natural Language Processing (ICON 2022)

  21. An Analysis of Abstractive Text Summarization Using Pre-trained Models

    Authors: Tohida Rehman, Suchandan Das, Debarshi Kumar Sanyal, Samiran Chattopadhyay

    Abstract: People nowadays use search engines like Google, Yahoo, and Bing to find information on the Internet. Due to explosion in data, it is helpful for users if they are provided relevant summaries of the search results rather than just links to webpages. Text summarization has become a vital approach to help consumers swiftly grasp vast amounts of information.In this paper, different pre-trained models… ▽ More

    Submitted 25 February, 2023; originally announced March 2023.

    Comments: 11 Pages, 6 Figures, 3 Tables

    Journal ref: https://link.springer.com/chapter/10.1007/978-981-19-1657-1_21(2022)

  22. arXiv:2303.12795  [pdf, other

    cs.CL cs.AI cs.LG

    Named Entity Recognition Based Automatic Generation of Research Highlights

    Authors: Tohida Rehman, Debarshi Kumar Sanyal, Prasenjit Majumder, Samiran Chattopadhyay

    Abstract: A scientific paper is traditionally prefaced by an abstract that summarizes the paper. Recently, research highlights that focus on the main findings of the paper have emerged as a complementary summary in addition to an abstract. However, highlights are not yet as common as abstracts, and are absent in many papers. In this paper, we aim to automatically generate research highlights using different… ▽ More

    Submitted 25 February, 2023; originally announced March 2023.

    Comments: 7 Pages, 3 Figures, 2 Tables

    Journal ref: https://aclanthology.org/2022.sdp-1.18

  23. Abstractive Text Summarization using Attentive GRU based Encoder-Decoder

    Authors: Tohida Rehman, Suchandan Das, Debarshi Kumar Sanyal, Samiran Chattopadhyay

    Abstract: In todays era huge volume of information exists everywhere. Therefore, it is very crucial to evaluate that information and extract useful, and often summarized, information out of it so that it may be used for relevant purposes. This extraction can be achieved through a crucial technique of artificial intelligence, namely, machine learning. Indeed automatic text summarization has emerged as an imp… ▽ More

    Submitted 25 February, 2023; originally announced February 2023.

    Comments: 9 pages, 2 Tables, 5 Figures

    Journal ref: https://link.springer.com/chapter/10.1007/978-981-19-4831-2_56(2022)

  24. arXiv:2302.09425  [pdf, other

    cs.AI

    A Neurodiversity-Inspired Solver for the Abstraction \& Reasoning Corpus (ARC) Using Visual Imagery and Program Synthesis

    Authors: James Ainooson, Deepayan Sanyal, Joel P. Michelson, Yuan Yang, Maithilee Kunda

    Abstract: Core knowledge about physical objects -- e.g., their permanency, spatial transformations, and interactions -- is one of the most fundamental building blocks of biological intelligence across humans and non-human animals. While AI techniques in certain domains (e.g. vision, NLP) have advanced dramatically in recent years, no current AI systems can yet match human abilities in flexibly applying core… ▽ More

    Submitted 31 October, 2023; v1 submitted 18 February, 2023; originally announced February 2023.

  25. Generation of Highlights from Research Papers Using Pointer-Generator Networks and SciBERT Embeddings

    Authors: Tohida Rehman, Debarshi Kumar Sanyal, Samiran Chattopadhyay, Plaban Kumar Bhowmick, Partha Pratim Das

    Abstract: Nowadays many research articles are prefaced with research highlights to summarize the main findings of the paper. Highlights not only help researchers precisely and quickly identify the contributions of a paper, they also enhance the discoverability of the article via search engines. We aim to automatically construct research highlights given certain segments of a research paper. We use a pointer… ▽ More

    Submitted 17 September, 2023; v1 submitted 14 February, 2023; originally announced February 2023.

    Comments: 19 Pages, 7 Figures, 8 Tables

    Journal ref: IEEE Access, 2023

  26. arXiv:2302.07137  [pdf, other

    cs.CV cs.AI

    Deep Non-Monotonic Reasoning for Visual Abstract Reasoning Tasks

    Authors: Yuan Yang, Deepayan Sanyal, Joel Michelson, James Ainooson, Maithilee Kunda

    Abstract: While achieving unmatched performance on many well-defined tasks, deep learning models have also been used to solve visual abstract reasoning tasks, which are relatively less well-defined, and have been widely used to measure human intelligence. However, current deep models struggle to match human abilities to solve such tasks with minimum data but maximum generalization. One limitation is that cu… ▽ More

    Submitted 8 February, 2023; originally announced February 2023.

  27. arXiv:2204.11428  [pdf, other

    cs.IR cs.HC

    Personal Research Knowledge Graphs

    Authors: Prantika Chakraborty, Sudakshina Dutta, Debarshi Kumar Sanyal

    Abstract: Maintaining research-related information in an organized manner can be challenging for a researcher. In this paper, we envision personal research knowledge graphs (PRKGs) as a means to represent structured information about the research activities of a researcher. PRKGs can be used to power intelligent personal assistants, and personalize various applications. We explore what entities and relation… ▽ More

    Submitted 25 April, 2022; originally announced April 2022.

  28. arXiv:2201.08450  [pdf, other

    cs.AI

    Automatic Item Generation of Figural Analogy Problems: A Review and Outlook

    Authors: Yuan Yang, Deepayan Sanyal, Joel Michelson, James Ainooson, Maithilee Kunda

    Abstract: Figural analogy problems have long been a widely used format in human intelligence tests. In the past four decades, more and more research has investigated automatic item generation for figural analogy problems, i.e., algorithmic approaches for systematically and automatically creating such problems. In cognitive science and psychometrics, this research can deepen our understandings of human analo… ▽ More

    Submitted 20 January, 2022; originally announced January 2022.

    Comments: Presented at The Ninth Advances in Cognitive Systems (ACS) Conference 2021 (arXiv:2201.06134)

    Report number: ACS2021/02

  29. Segmenting Scientific Abstracts into Discourse Categories: A Deep Learning-Based Approach for Sparse Labeled Data

    Authors: Soumya Banerjee, Debarshi Kumar Sanyal, Samiran Chattopadhyay, Plaban Kumar Bhowmick, Parthapratim Das

    Abstract: The abstract of a scientific paper distills the contents of the paper into a short paragraph. In the biomedical literature, it is customary to structure an abstract into discourse categories like BACKGROUND, OBJECTIVE, METHOD, RESULT, and CONCLUSION, but this segmentation is uncommon in other fields like computer science. Explicit categories could be helpful for more granular, that is, discourse-l… ▽ More

    Submitted 27 May, 2020; v1 submitted 11 May, 2020; originally announced May 2020.

    Comments: to appear in the proceedings of JCDL'2020

    ACM Class: I.5.1; H.3.7

  30. arXiv:2002.03131  [pdf, other

    cs.CV

    Variable-Viewpoint Representations for 3D Object Recognition

    Authors: Tengyu Ma, Joel Michelson, James Ainooson, Deepayan Sanyal, Xiaohan Wang, Maithilee Kunda

    Abstract: For the problem of 3D object recognition, researchers using deep learning methods have developed several very different input representations, including "multi-view" snapshots taken from discrete viewpoints around an object, as well as "spherical" representations consisting of a dense map of essentially ray-traced samples of the object from all directions. These representations offer trade-offs in… ▽ More

    Submitted 8 February, 2020; originally announced February 2020.

    Comments: 8 pages, 6 figures

  31. arXiv:1306.3903  [pdf

    cs.NI

    Designing an Efficient Delay Sensitive Routing Metric for IEEE 802.16 Mesh Networks

    Authors: Ishita Bhakta, Sandip Chakraborty, Barsha Mitra, Debarshi Kumar Sanyal, Samiran Chattopadhyay, Matangini Chattopadhyay

    Abstract: Quality of Service provisioning is one of the major design goals of IEEE 802.16 mesh networks. In order to provide quality delivery of delay sensitive services such as voice, video etc., it is required to route such traffic over a minimum delay path. In this paper we propose a routing metric for delay sensitive services in IEEE 802.16 mesh networks. We design a new cross layer routing metric, name… ▽ More

    Submitted 15 July, 2013; v1 submitted 17 June, 2013; originally announced June 2013.

    Comments: This paper has been presented at the International Conference on Wireless and Optical Communications, May 2011, China