Computation and Language

Authors and titles for August 2024

Total of 1234 entries : 1-25 76-100 101-125 126-150 151-175 176-200 201-225 226-250 ... 1226-1234

Showing up to 25 entries per page: fewer | more | all

[151] arXiv:2408.03505 [pdf, html, other]: Title: Optimus: Accelerating Large-Scale Multi-Modal LLM Training by Bubble Exploitation

Weiqi Feng, Yangrui Chen, Shaoyu Wang, Yanghua Peng, Haibin Lin, Minlan Yu

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Distributed, Parallel, and Cluster Computing (cs.DC)
[152] arXiv:2408.03506 [pdf, html, other]: Title: 1.5-Pints Technical Report: Pretraining in Days, Not Months -- Your Language Model Thrives on Quality Data

Calvin Tan, Jerome Wang

Comments: Technical Report for 1.5-Pints

Subjects: Computation and Language (cs.CL)
[153] arXiv:2408.03524 [pdf, html, other]: Title: EgyBERT: A Large Language Model Pretrained on Egyptian Dialect Corpora

Faisal Qarah

Subjects: Computation and Language (cs.CL)
[154] arXiv:2408.03541 [pdf, html, other]: Title: EXAONE 3.0 7.8B Instruction Tuned Language Model

LG AI Research: Soyoung An, Kyunghoon Bae, Eunbi Choi, Stanley Jungkyu Choi, Yemuk Choi, Seokhee Hong, Yeonjung Hong, Junwon Hwang, Hyojin Jeon, Gerrard Jeongwon Jo, Hyunjik Jo, Jiyeon Jung, Yountae Jung, Euisoon Kim, Hyosang Kim, Joonkee Kim, Seonghwan Kim, Soyeon Kim, Sunkyoung Kim, Yireun Kim, Youchul Kim, Edward Hwayoung Lee, Haeju Lee, Honglak Lee, Jinsik Lee, Kyungmin Lee, Moontae Lee, Seungjun Lee, Woohyung Lim, Sangha Park, Sooyoun Park, Yongmin Park, Boseong Seo, Sihoon Yang, Heuiyeen Yeen, Kyungjae Yoo, Hyeongu Yun

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[155] arXiv:2408.03544 [pdf, other]: Title: NatLan: Native Language Prompting Facilitates Knowledge Elicitation Through Language Trigger Provision and Domain Trigger Retention

Baixuan Li, Yunlong Fan, Tianyi Ma, Zhiqiang Gao

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[156] arXiv:2408.03554 [pdf, html, other]: Title: Empirical Analysis of Large Vision-Language Models against Goal Hijacking via Visual Prompt Injection

Subaru Kimura, Ryota Tanaka, Shumpei Miyawaki, Jun Suzuki, Keisuke Sakaguchi

Comments: 8 pages, 6 figures, Accepted to NAACL 2024 SRW

Subjects: Computation and Language (cs.CL); Cryptography and Security (cs.CR); Machine Learning (cs.LG)
[157] arXiv:2408.03562 [pdf, html, other]: Title: A Comparison of LLM Finetuning Methods & Evaluation Metrics with Travel Chatbot Use Case

Sonia Meyer, Shreya Singh, Bertha Tam, Christopher Ton, Angel Ren

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[158] arXiv:2408.03617 [pdf, html, other]: Title: Is Child-Directed Speech Effective Training Data for Language Models?

Steven Y. Feng, Noah D. Goodman, Michael C. Frank

Comments: EMNLP 2024. Code and data at this https URL

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[159] arXiv:2408.03618 [pdf, other]: Title: A Logical Fallacy-Informed Framework for Argument Generation

Luca Mouchel, Debjit Paul, Shaobo Cui, Robert West, Antoine Bosselut, Boi Faltings

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[160] arXiv:2408.03622 [pdf, other]: Title: Improving the quality of Persian clinical text with a novel spelling correction system

Seyed Mohammad Sadegh Dashti, Seyedeh Fatemeh Dashti

Journal-ref: BMC Med Inform Decis Mak 24, 220 (2024)

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[161] arXiv:2408.03630 [pdf, html, other]: Title: PAGED: A Benchmark for Procedural Graphs Extraction from Documents

Weihong Du, Wenrui Liao, Hongru Liang, Wenqiang Lei

Comments: Accepted to The 62nd Annual Meeting of the Association for Computational Linguistics (ACL 2024)

Subjects: Computation and Language (cs.CL)
[162] arXiv:2408.03633 [pdf, html, other]: Title: CARE: A Clue-guided Assistant for CSRs to Read User Manuals

Weihong Du, Jia Liu, Zujie Wen, Dingnan Jin, Hongru Liang, Wenqiang Lei

Comments: Accepted to The 62nd Annual Meeting of the Association for Computational Linguistics (ACL 2024)

Subjects: Computation and Language (cs.CL)
[163] arXiv:2408.03652 [pdf, html, other]: Title: mucAI at WojoodNER 2024: Arabic Named Entity Recognition with Nearest Neighbor Search

Ahmed Abdou, Tasneem Mohsen

Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[164] arXiv:2408.03675 [pdf, html, other]: Title: NACL: A General and Effective KV Cache Eviction Framework for LLMs at Inference Time

Yilong Chen, Guoxia Wang, Junyuan Shang, Shiyao Cui, Zhenyu Zhang, Tingwen Liu, Shuohuan Wang, Yu Sun, Dianhai Yu, Hua Wu

Comments: Accepted by ACL 2024 (main conference, long paper)

Subjects: Computation and Language (cs.CL)
[165] arXiv:2408.03706 [pdf, html, other]: Title: Local Topology Measures of Contextual Language Model Latent Spaces With Applications to Dialogue Term Extraction

Benjamin Matthias Ruppik, Michael Heck, Carel van Niekerk, Renato Vukovic, Hsien-chin Lin, Shutong Feng, Marcus Zibrowius, Milica Gašić

Comments: Accepted as a long paper to SIGDIAL 2024. 9 pages, 2 figures, 3 tables

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[166] arXiv:2408.03732 [pdf, html, other]: Title: Question Rephrasing for Quantifying Uncertainty in Large Language Models: Applications in Molecular Chemistry Tasks

Zizhang Chen, Pengyu Hong, Sandeep Madireddy

Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG); Quantitative Methods (q-bio.QM)
[167] arXiv:2408.03762 [pdf, html, other]: Title: 'Finance Wizard' at the FinLLM Challenge Task: Financial Text Summarization

Meisin Lee, Soon Lay-Ki

Subjects: Computation and Language (cs.CL)
[168] arXiv:2408.03811 [pdf, html, other]: Title: Generative Language Models with Retrieval Augmented Generation for Automated Short Answer Scoring

Zifan Wang, Christopher Ormerod

Comments: 20 pages, 2 figures

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Information Retrieval (cs.IR)
[169] arXiv:2408.03837 [pdf, html, other]: Title: WalledEval: A Comprehensive Safety Evaluation Toolkit for Large Language Models

Prannaya Gupta, Le Qi Yau, Hao Han Low, I-Shiang Lee, Hugo Maximus Lim, Yu Xin Teoh, Jia Hng Koh, Dar Win Liew, Rishabh Bhardwaj, Rajat Bhardwaj, Soujanya Poria

Comments: Under review

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[170] arXiv:2408.03849 [pdf, html, other]: Title: Hate Speech Detection and Classification in Amharic Text with Deep Learning

Samuel Minale Gashe, Seid Muhie Yimam, Yaregal Assabie

Comments: Dataset: this https URL

Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[171] arXiv:2408.03855 [pdf, html, other]: Title: Why transformers are obviously good models of language

Felix Hill

Subjects: Computation and Language (cs.CL)
[172] arXiv:2408.03871 [pdf, html, other]: Title: Large Language Models for Biomedical Text Simplification: Promising But Not There Yet

Zihao Li, Samuel Belkadi, Nicolo Micheletti, Lifeng Han, Matthew Shardlow, Goran Nenadic

Comments: Extended system report for PLABA-2023. arXiv admin note: substantial text overlap with arXiv:2309.13202

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[173] arXiv:2408.03874 [pdf, html, other]: Title: Personalized Clinical Note Generation from Doctor-Patient Conversations

Nathan Brake, Thomas Schaaf

Subjects: Computation and Language (cs.CL)
[174] arXiv:2408.03899 [pdf, html, other]: Title: Simplifying Scholarly Abstracts for Accessible Digital Libraries

Haining Wang, Jason Clark

Comments: Initial submission to JCDL2024

Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computers and Society (cs.CY); Digital Libraries (cs.DL)
[175] arXiv:2408.03900 [pdf, html, other]: Title: Speech-MASSIVE: A Multilingual Speech Dataset for SLU and Beyond

Beomseok Lee, Ioan Calapodescu, Marco Gaido, Matteo Negri, Laurent Besacier

Comments: Accepted at INTERSPEECH 2024. This version includes the same content but with additional appendices

Subjects: Computation and Language (cs.CL); Sound (cs.SD); Audio and Speech Processing (eess.AS)

Total of 1234 entries : 1-25 76-100 101-125 126-150 151-175 176-200 201-225 226-250 ... 1226-1234

Showing up to 25 entries per page: fewer | more | all