We gratefully acknowledge support from
the Simons Foundation and member institutions.

Computation and Language

Authors and titles for recent submissions

[ total of 352 entries: 1-98 | 99-196 | 197-294 | 295-352 ]
[ showing 98 entries per page: fewer | more | all ]

Tue, 16 Apr 2024 (showing first 98 of 137 entries)

[1]  arXiv:2404.09982 [pdf, other]
Title: Memory Sharing for Large Language Model based Agents
Subjects: Computation and Language (cs.CL)
[2]  arXiv:2404.09980 [pdf, other]
Title: Context Does Matter: Implications for Crowdsourced Evaluation Labels in Task-Oriented Dialogue Systems
Comments: Accepted at NAACL 2024 Findings
Subjects: Computation and Language (cs.CL); Human-Computer Interaction (cs.HC); Information Retrieval (cs.IR)
[3]  arXiv:2404.09971 [pdf, other]
Title: Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs
Subjects: Computation and Language (cs.CL)
[4]  arXiv:2404.09937 [pdf, other]
Title: Compression Represents Intelligence Linearly
Comments: Preprint. Data and code are available at this https URL
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Information Theory (cs.IT); Machine Learning (cs.LG)
[5]  arXiv:2404.09911 [pdf, other]
Title: ChatShop: Interactive Information Seeking with Language Agents
Subjects: Computation and Language (cs.CL)
[6]  arXiv:2404.09894 [pdf, ps, other]
Title: Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
Subjects: Computation and Language (cs.CL); Software Engineering (cs.SE)
[7]  arXiv:2404.09830 [pdf, other]
Title: Negation Triplet Extraction with Syntactic Dependency and Semantic Consistency
Comments: Accepted by COLING 2024
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[8]  arXiv:2404.09824 [pdf, other]
Title: Impact of Preference Noise on the Alignment Performance of Generative Language Models
Subjects: Computation and Language (cs.CL)
[9]  arXiv:2404.09785 [pdf, other]
Title: Benchmarking Llama2, Mistral, Gemma and GPT for Factuality, Toxicity, Bias and Propensity for Hallucinations
Comments: 14 pages, 8 figures, 18 tables
Subjects: Computation and Language (cs.CL)
[10]  arXiv:2404.09763 [pdf, other]
Title: KG-CTG: Citation Generation through Knowledge Graph-guided Large Language Models
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[11]  arXiv:2404.09754 [pdf, other]
Title: Resilience of Large Language Models for Noisy Instructions
Comments: 12 pages
Subjects: Computation and Language (cs.CL)
[12]  arXiv:2404.09753 [pdf, other]
Title: Personalized Collaborative Fine-Tuning for On-Device Large Language Models
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[13]  arXiv:2404.09717 [pdf, other]
Title: Unveiling Imitation Learning: Exploring the Impact of Data Falsity to Large Language Model
Authors: Hyunsoo Cho
Comments: Under review @ *ACL
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[14]  arXiv:2404.09696 [pdf, other]
Title: Are Large Language Models Reliable Argument Quality Annotators?
Comments: 18 pages, 5 figures, 5 tables
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Emerging Technologies (cs.ET)
[15]  arXiv:2404.09682 [pdf, other]
Title: Multi-News+: Cost-efficient Dataset Cleansing via LLM-based Data Annotation
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[16]  arXiv:2404.09615 [pdf, other]
Title: If there's a Trigger Warning, then where's the Trigger? Investigating Trigger Warnings at the Passage Level
Subjects: Computation and Language (cs.CL); Computers and Society (cs.CY)
[17]  arXiv:2404.09593 [pdf, other]
Title: Improving Recall of Large Language Models: A Model Collaboration Approach for Relational Triple Extraction
Comments: Accepted at LREC-COLING 2024 main conference
Subjects: Computation and Language (cs.CL)
[18]  arXiv:2404.09579 [pdf, ps, other]
Title: Modelling Language
Authors: Jumbly Grindrod
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[19]  arXiv:2404.09577 [pdf, ps, other]
Title: Transformers, Contextualism, and Polysemy
Authors: Jumbly Grindrod
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[20]  arXiv:2404.09576 [pdf, ps, other]
Title: Large language models and linguistic intentionality
Authors: Jumbly Grindrod
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[21]  arXiv:2404.09565 [pdf, other]
Title: Reliability Estimation of News Media Sources: Birds of a Feather Flock Together
Comments: Accepted to NAACL 2024 Main Conference
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computers and Society (cs.CY); Machine Learning (cs.LG)
[22]  arXiv:2404.09492 [pdf, other]
Title: Bridging the Gap between Different Vocabularies for LLM Ensemble
Comments: Accepted to the main conference of NAACL 2024
Subjects: Computation and Language (cs.CL)
[23]  arXiv:2404.09486 [pdf, other]
Title: MMCode: Evaluating Multi-Modal Code Large Language Models with Visually Rich Programming Problems
Comments: 46 pages, 21 figures and 6 tables
Subjects: Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV); Software Engineering (cs.SE)
[24]  arXiv:2404.09480 [pdf, other]
Title: Mitigating Hallucination in Abstractive Summarization with Domain-Conditional Mutual Information
Comments: Accepted by Findings of NAACL 2024
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[25]  arXiv:2404.09416 [pdf, ps, other]
Title: Automatic Knowledge Graph Construction for Judicial Cases
Subjects: Computation and Language (cs.CL)
[26]  arXiv:2404.09405 [pdf, other]
Title: Few-shot Name Entity Recognition on StackOverflow
Comments: 5 pages
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[27]  arXiv:2404.09383 [pdf, other]
Title: Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
Comments: IJCNLP 2017
Subjects: Computation and Language (cs.CL)
[28]  arXiv:2404.09371 [pdf, other]
Title: The Effect of Data Partitioning Strategy on Model Generalizability: A Case Study of Morphological Segmentation
Comments: Accepted to 2024 Annual Conference of the North American Chapter of the Association for Computational Linguistics (16 pages including 9 tables and 1 figure)
Subjects: Computation and Language (cs.CL)
[29]  arXiv:2404.09366 [pdf, ps, other]
Title: Understanding the Role of Temperature in Diverse Question Generation by GPT-4
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[30]  arXiv:2404.09339 [pdf, other]
Title: Towards Practical Tool Usage for Continually Learning LLMs
Comments: 20 pages, 11 tables, 7 figures
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[31]  arXiv:2404.09338 [pdf, other]
Title: Entropy Guided Extrapolative Decoding to Improve Factuality in Large Language Models
Comments: Work in Progress
Subjects: Computation and Language (cs.CL)
[32]  arXiv:2404.09336 [pdf, other]
Title: Self-Selected Attention Span for Accelerating Large Language Model Inference
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[33]  arXiv:2404.09329 [pdf, ps, other]
Title: Large Language Models are as persuasive as humans, but why? About the cognitive effort and moral-emotional language of LLM arguments
Subjects: Computation and Language (cs.CL)
[34]  arXiv:2404.09299 [pdf, other]
Title: Reap the Wild Wind: Detecting Media Storms in Large-Scale News Corpora
Subjects: Computation and Language (cs.CL)
[35]  arXiv:2404.09296 [pdf, other]
Title: Cross-Data Knowledge Graph Construction for LLM-enabled Educational Question-Answering System: A~Case~Study~at~HCMUT
Comments: 8 pages, 7 figures
Subjects: Computation and Language (cs.CL)
[36]  arXiv:2404.09260 [pdf, other]
Title: JaFIn: Japanese Financial Instruction Dataset
Comments: 10 pages, 1 figure
Subjects: Computation and Language (cs.CL); Computational Engineering, Finance, and Science (cs.CE)
[37]  arXiv:2404.09221 [pdf, other]
Title: Towards Fast Inference: Exploring and Improving Blockwise Parallel Drafts
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[38]  arXiv:2404.09220 [pdf, other]
Title: Compass: Large Multilingual Language Model for South-east Asia
Authors: Sophia Maria
Subjects: Computation and Language (cs.CL)
[39]  arXiv:2404.09206 [pdf, other]
Title: DKE-Research at SemEval-2024 Task 2: Incorporating Data Augmentation with Generative Models and Biomedical Knowledge to Enhance Inference Robustness
Subjects: Computation and Language (cs.CL)
[40]  arXiv:2404.09170 [pdf, other]
Title: Post-Semantic-Thinking: A Robust Strategy to Distill Reasoning Capacity from Large Language Models
Subjects: Computation and Language (cs.CL)
[41]  arXiv:2404.09163 [pdf, other]
Title: GeMQuAD : Generating Multilingual Question Answering Datasets from Large Language Models using Few Shot Learning
Comments: Accepted to The 37th International Conference on Neural Information Processing Systems (NeurIPS 2023)December 10-16, 2023 - SyntheticData4ML workshop, New Orleans, United States this https URL
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[42]  arXiv:2404.09145 [pdf, other]
Title: ToNER: Type-oriented Named Entity Recognition with Generative Language Model
Comments: Accepted at LREC-COLING 2024
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[43]  arXiv:2404.09138 [pdf, other]
Title: From Bytes to Borsch: Fine-Tuning Gemma and Mistral for the Ukrainian Language Representation
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[44]  arXiv:2404.09136 [pdf, other]
Title: TLDR at SemEval-2024 Task 2: T5-generated clinical-Language summaries for DeBERTa Report Analysis
Journal-ref: In Proceedings of the 18th International Workshop on Semantic Evaluation (SemEval-2024), pages 507-516, Mexico City, Mexico. Association for Computational Linguistics
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[45]  arXiv:2404.09135 [pdf, ps, other]
Title: Unveiling LLM Evaluation Focused on Metrics: Challenges and Solutions
Subjects: Computation and Language (cs.CL)
[46]  arXiv:2404.09129 [pdf, other]
Title: When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models
Comments: NAACL 2024 Findings paper (Camera-Ready Version)
Subjects: Computation and Language (cs.CL)
[47]  arXiv:2404.09127 [pdf, other]
Title: Confidence Calibration and Rationalization for LLMs via Multi-Agent Deliberation
Comments: Accepted at ICLR 2024 Workshop on Reliable and Responsible Foundation Models
Subjects: Computation and Language (cs.CL)
[48]  arXiv:2404.09077 [pdf, other]
Title: CuriousLLM: Elevating Multi-Document QA with Reasoning-Infused Knowledge Graph Prompting
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Information Retrieval (cs.IR); Machine Learning (cs.LG)
[49]  arXiv:2404.09047 [pdf, other]
Title: Multilingual Evaluation of Semantic Textual Relatedness
Comments: 8 pages
Subjects: Computation and Language (cs.CL)
[50]  arXiv:2404.09045 [pdf, other]
Title: Adapting Mental Health Prediction Tasks for Cross-lingual Learning via Meta-Training and In-context Learning with Large Language Model
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computers and Society (cs.CY); Machine Learning (cs.LG)
[51]  arXiv:2404.09043 [pdf, other]
Title: Do LLMs Play Dice? Exploring Probability Distribution Sampling in Large Language Models for Behavioral Simulation
Subjects: Computation and Language (cs.CL)
[52]  arXiv:2404.09027 [pdf, other]
Title: MING-MOE: Enhancing Medical Multi-Task Learning in Large Language Models with Sparse Mixture of Low-Rank Adapter Experts
Comments: 15 pages, 3 figures
Subjects: Computation and Language (cs.CL)
[53]  arXiv:2404.09002 [pdf, other]
Title: WikiSplit++: Easy Data Refinement for Split and Rephrase
Comments: Accepted at LREC-COLING 2024
Subjects: Computation and Language (cs.CL)
[54]  arXiv:2404.08997 [pdf, other]
Title: Labeled Morphological Segmentation with Semi-Markov Models
Comments: CoNLL 2015
Subjects: Computation and Language (cs.CL)
[55]  arXiv:2404.08977 [pdf, other]
Title: RoNID: New Intent Discovery with Generated-Reliable Labels and Cluster-friendly Representations
Comments: DASFAA 2024
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[56]  arXiv:2404.08974 [pdf, other]
Title: OOVs in the Spotlight: How to Inflect them?
Comments: To be published in LREC-COLING 2024. 12 pages, 3 figures
Subjects: Computation and Language (cs.CL)
[57]  arXiv:2404.08949 [pdf, other]
Title: Multimodal Cross-Document Event Coreference Resolution Using Linear Semantic Transfer and Mixed-Modality Ensembles
Comments: To appear at LREC-COLING 2024
Subjects: Computation and Language (cs.CL)
[58]  arXiv:2404.08938 [pdf, other]
Title: Enforcing Paraphrase Generation via Controllable Latent Diffusion
Subjects: Computation and Language (cs.CL)
[59]  arXiv:2404.08888 [pdf, other]
Title: Towards Enhancing Health Coaching Dialogue in Low-Resource Settings
Comments: Accepted to the main conference of COLING 2022
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[60]  arXiv:2404.08865 [pdf, other]
Title: LLM In-Context Recall is Prompt Dependent
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[61]  arXiv:2404.08856 [pdf, other]
Title: On Speculative Decoding for Multimodal Large Language Models
Comments: Accepted as a spotlight paper to ELVM workshop at CVPR 2024
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[62]  arXiv:2404.08836 [pdf, other]
Title: BERT-LSH: Reducing Absolute Compute For Attention
Comments: 10 pages, 5 figures
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[63]  arXiv:2404.08821 [pdf, other]
Title: Constrained C-Test Generation via Mixed-Integer Programming
Comments: Github: this https URL
Subjects: Computation and Language (cs.CL)
[64]  arXiv:2404.08817 [pdf, other]
Title: Revisiting Code Similarity Evaluation with Abstract Syntax Tree Edit Distance
Subjects: Computation and Language (cs.CL); Programming Languages (cs.PL); Software Engineering (cs.SE)
[65]  arXiv:2404.08816 [pdf, other]
Title: Evaluating the Quality of Answers in Political Q&A Sessions with Large Language Models
Subjects: Computation and Language (cs.CL); Econometrics (econ.EM)
[66]  arXiv:2404.08806 [pdf, other]
Title: CreativEval: Evaluating Creativity of LLM-Based Hardware Code Generation
Subjects: Computation and Language (cs.CL)
[67]  arXiv:2404.08760 [pdf, other]
Title: The Generation Gap:Exploring Age Bias in Large Language Models
Comments: 4 pages
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[68]  arXiv:2404.08705 [pdf, other]
Title: Introducing L2M3, A Multilingual Medical Large Language Model to Advance Health Equity in Low-Resource Regions
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[69]  arXiv:2404.08704 [pdf, other]
Title: MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[70]  arXiv:2404.08700 [pdf, other]
Title: Is Your LLM Outdated? Benchmarking LLMs & Alignment Algorithms for Time-Sensitive Knowledge
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[71]  arXiv:2404.08699 [pdf, other]
Title: Analyzing the Impact of Data Selection and Fine-Tuning on Economic and Political Biases in LLMs
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[72]  arXiv:2404.08698 [pdf, other]
Title: Lossless Acceleration of Large Language Model via Adaptive N-gram Parallel Decoding
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[73]  arXiv:2404.08695 [pdf, other]
Title: Enhancing Question Answering for Enterprise Knowledge Bases using Large Language Models
Comments: DASFAA 2024 Accepted
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Information Retrieval (cs.IR)
[74]  arXiv:2404.08690 [pdf, other]
Title: Towards Building a Robust Toxicity Predictor
Comments: ACL 2023 /
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Cryptography and Security (cs.CR); Machine Learning (cs.LG)
[75]  arXiv:2404.08686 [pdf, other]
Title: Extractive text summarisation of Privacy Policy documents using machine learning approaches
Authors: Chanwoo Choi
Comments: University of Edinburgh MInf (Master of Informatics) Thesis, 52 pages, 13 figures, Submitted and approved by the institution in May 2022
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[76]  arXiv:2404.08685 [pdf, ps, other]
Title: Neural Sequence-to-Sequence Modeling with Attention by Leveraging Deep Learning Architectures for Enhanced Contextual Understanding in Abstractive Text Summarization
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[77]  arXiv:2404.08684 [pdf, ps, other]
Title: Is English the New Programming Language? How About Pseudo-code Engineering?
Journal-ref: Acta Sci. (Canoas), 26(1), 157-204, Jan./Feb. 2024
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Human-Computer Interaction (cs.HC)
[78]  arXiv:2404.08683 [pdf, other]
Title: Text clustering applied to data augmentation in legal contexts
Comments: 23 pages, 4 figures. submitted to Artificial Intelligence and Law Journal
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[79]  arXiv:2404.08681 [pdf, other]
Title: EFSA: Towards Event-Level Financial Sentiment Analysis
Subjects: Computation and Language (cs.CL)
[80]  arXiv:2404.08680 [pdf, other]
Title: Automating Research Synthesis with Domain-Specific Large Language Model Fine-Tuning
Subjects: Computation and Language (cs.CL); Digital Libraries (cs.DL); Information Retrieval (cs.IR)
[81]  arXiv:2404.08679 [pdf, other]
Title: Your Finetuned Large Language Model is Already a Powerful Out-of-distribution Detector
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Machine Learning (stat.ML)
[82]  arXiv:2404.08676 [pdf, other]
Title: ALERT: A Comprehensive Benchmark for Assessing Large Language Models' Safety through Red Teaming
Comments: 19 pages, preprint
Subjects: Computation and Language (cs.CL); Computers and Society (cs.CY); Machine Learning (cs.LG)
[83]  arXiv:2404.08674 [pdf, other]
Title: Effects of Different Prompts on the Quality of GPT-4 Responses to Dementia Care Questions
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Human-Computer Interaction (cs.HC)
[84]  arXiv:2404.08673 [pdf, other]
Title: Sentiment analysis and random forest to classify LLM versus human source applied to Scientific Texts
Comments: 12 Pages, 3 tables, 6 figures
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Computers and Society (cs.CY); Machine Learning (cs.LG)
[85]  arXiv:2404.08666 [pdf, other]
Title: Revealing Trends in Datasets from the 2022 ACL and EMNLP Conferences
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[86]  arXiv:2404.08661 [pdf, other]
Title: The Comparison of Translationese in Machine Translation and Human Transation in terms of Translation Relations
Authors: Fan Zhou
Subjects: Computation and Language (cs.CL)
[87]  arXiv:2404.08656 [pdf, other]
Title: Linear Cross-document Event Coreference Resolution with X-AMR
Comments: LREC-COLING 2024 main conference
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[88]  arXiv:2404.08655 [pdf, other]
Title: Transformer-based Joint Modelling for Automatic Essay Scoring and Off-Topic Detection
Comments: Accepted in LREC-COLING 2024
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[89]  arXiv:2404.08654 [pdf, ps, other]
Title: Optimal path for Biomedical Text Summarization Using Pointer GPT
Comments: 3 pages, 3 figures
Journal-ref: KSC2023
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[90]  arXiv:2404.09992 (cross-list from cs.CV) [pdf, other]
Title: MMInA: Benchmarking Multihop Multimodal Internet Agents
Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[91]  arXiv:2404.09956 (cross-list from cs.SD) [pdf, other]
Title: Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
Comments: this https URL
Subjects: Sound (cs.SD); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Audio and Speech Processing (eess.AS)
[92]  arXiv:2404.09932 (cross-list from cs.LG) [pdf, other]
[93]  arXiv:2404.09897 (cross-list from cs.AI) [pdf, other]
Title: Progressive Knowledge Graph Completion
Comments: 14 pages, 10 figures
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG)
[94]  arXiv:2404.09889 (cross-list from cs.IR) [pdf, other]
Title: Is Table Retrieval a Solved Problem? Join-Aware Multi-Table Retrieval
Subjects: Information Retrieval (cs.IR); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[95]  arXiv:2404.09868 (cross-list from cs.SE) [pdf, other]
Title: AI-Driven Statutory Reasoning via Software Engineering Methods
Authors: Rohan Padhye
Subjects: Software Engineering (cs.SE); Computation and Language (cs.CL); Computers and Society (cs.CY)
[96]  arXiv:2404.09841 (cross-list from eess.AS) [pdf, other]
Title: Anatomy of Industrial Scale Multilingual ASR
Subjects: Audio and Speech Processing (eess.AS); Computation and Language (cs.CL); Machine Learning (cs.LG); Sound (cs.SD)
[97]  arXiv:2404.09737 (cross-list from cs.LG) [pdf, other]
Title: Quantization of Large Language Models with an Overdetermined Basis
Subjects: Machine Learning (cs.LG); Computation and Language (cs.CL)
[98]  arXiv:2404.09695 (cross-list from cs.LG) [pdf, other]
Title: LoRAP: Transformer Sub-Layers Deserve Differentiated Structured Compression for Large Language Models
Comments: 8 pages,4 figures
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[ total of 352 entries: 1-98 | 99-196 | 197-294 | 295-352 ]
[ showing 98 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, cs, new, 2404, contact, help  (Access key information)