We gratefully acknowledge support from
the Simons Foundation and member institutions.

Computation and Language

Authors and titles for recent submissions, skipping first 168

[ total of 352 entries: 1-25 | ... | 94-118 | 119-143 | 144-168 | 169-193 | 194-218 | 219-243 | 244-268 | ... | 344-352 ]
[ showing 25 entries per page: fewer | more | all ]

Mon, 15 Apr 2024 (continued, showing last 15 of 46 entries)

[169]  arXiv:2404.08509 (cross-list from cs.DC) [pdf, other]
Title: Efficient Interactive LLM Serving with Proxy Model-based Sequence Length Prediction
Comments: Accepted at AIOps'24
Subjects: Distributed, Parallel, and Cluster Computing (cs.DC); Computation and Language (cs.CL); Machine Learning (cs.LG)
[170]  arXiv:2404.08495 (cross-list from cs.LG) [pdf, other]
Title: Dataset Reset Policy Optimization for RLHF
Comments: 28 pages, 6 tables, 3 Figures, 3 Algorithms
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[171]  arXiv:2404.08480 (cross-list from cs.LG) [pdf, other]
Title: Decoding AI: The inside story of data analysis in ChatGPT
Comments: 15 pages with figures and appendix
Subjects: Machine Learning (cs.LG); Computation and Language (cs.CL); Computation (stat.CO)
[172]  arXiv:2404.08417 (cross-list from cs.LG) [pdf, other]
Title: AdapterSwap: Continuous Training of LLMs with Data Removal and Access-Control Guarantees
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[173]  arXiv:2404.08309 (cross-list from cs.CR) [pdf, other]
Title: Subtoxic Questions: Dive Into Attitude Change of LLM's Response in Jailbreak Attempts
Comments: 4 pages, 2 figures. This paper was submitted to The 7th Deep Learning Security and Privacy Workshop (DLSP 2024) and was accepted as extended abstract, see this https URL
Subjects: Cryptography and Security (cs.CR); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[174]  arXiv:2404.08189 (cross-list from cs.LG) [pdf, other]
Title: Reducing hallucination in structured outputs via Retrieval-Augmented Generation
Comments: To be presented at NAACL 2024. 11 pages and 4 figures
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Information Retrieval (cs.IR)
[175]  arXiv:2404.08164 (cross-list from stat.ML) [pdf, other]
Title: Language Model Prompt Selection via Simulation Optimization
Subjects: Machine Learning (stat.ML); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG)
[176]  arXiv:2404.08134 (cross-list from cs.IR) [pdf, other]
Title: Extending Translate-Train for ColBERT-X to African Language CLIR
Comments: 10 pages, 2 figures. System description paper for HLTCOE's participation in CIRAL@FIRE 2023
Subjects: Information Retrieval (cs.IR); Computation and Language (cs.CL)
[177]  arXiv:2404.08111 (cross-list from cs.CV) [pdf, other]
Title: S3Editor: A Sparse Semantic-Disentangled Self-Training Framework for Face Video Editing
Subjects: Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[178]  arXiv:2404.08080 (cross-list from cs.LG) [pdf, other]
Title: Variance-reduced Zeroth-Order Methods for Fine-Tuning Language Models
Comments: 29 pages, 25 tables, 9 figures
Subjects: Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Optimization and Control (math.OC)
[179]  arXiv:2404.08020 (cross-list from cs.AI) [pdf, other]
Title: Augmenting Knowledge Graph Hierarchies Using Neural Transformers
Comments: European Conference on Information Retrieval 2024
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Digital Libraries (cs.DL); Information Retrieval (cs.IR); Machine Learning (cs.LG)
[180]  arXiv:2404.08018 (cross-list from cs.SE) [pdf, other]
Title: Analyzing the Performance of Large Language Models on Code Summarization
Subjects: Software Engineering (cs.SE); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[181]  arXiv:2404.08008 (cross-list from cs.LG) [pdf, other]
Title: Sample-Efficient Human Evaluation of Large Language Models via Maximum Discrepancy Competition
Comments: 32 pages, 6 figures
Subjects: Machine Learning (cs.LG); Computation and Language (cs.CL); Human-Computer Interaction (cs.HC)
[182]  arXiv:2404.08001 (cross-list from hep-ph) [pdf, other]
Title: Xiwu: A Basis Flexible and Learnable LLM for High Energy Physics
Comments: 15 pages, 8 figures
Subjects: High Energy Physics - Phenomenology (hep-ph); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG); High Energy Physics - Experiment (hep-ex); Computational Physics (physics.comp-ph)
[183]  arXiv:2404.07999 (cross-list from cs.LG) [pdf, other]
Title: A Multi-Level Framework for Accelerating Training Transformer Models
Comments: ICLR 2024
Subjects: Machine Learning (cs.LG); Computation and Language (cs.CL)

Fri, 12 Apr 2024 (showing first 10 of 59 entries)

[184]  arXiv:2404.07982 [pdf, other]
Title: Language Imbalance Can Boost Cross-lingual Generalisation
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[185]  arXiv:2404.07979 [pdf, other]
Title: LLoCO: Learning Long Contexts Offline
Comments: The first two authors contributed equally to this work
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[186]  arXiv:2404.07965 [pdf, other]
Title: Rho-1: Not All Tokens Are What You Need
Comments: First two authors equal contribution
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[187]  arXiv:2404.07922 [pdf, other]
Title: LaVy: Vietnamese Multimodal Large Language Model
Comments: 4 pages
Subjects: Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[188]  arXiv:2404.07921 [pdf, other]
Title: AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs
Authors: Zeyi Liao, Huan Sun
Subjects: Computation and Language (cs.CL)
[189]  arXiv:2404.07904 [pdf, other]
Title: HGRN2: Gated Linear RNNs with State Expansion
Comments: Techinical Report. Yiran Zhong is the corresponding author. The source code is available at this https URL
Subjects: Computation and Language (cs.CL)
[190]  arXiv:2404.07900 [pdf, other]
Title: High-Dimension Human Value Representation in Large Language Models
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[191]  arXiv:2404.07879 [pdf, other]
Title: Analyzing Toxicity in Deep Conversations: A Reddit Case Study
Subjects: Computation and Language (cs.CL); Computers and Society (cs.CY); Social and Information Networks (cs.SI)
[192]  arXiv:2404.07851 [pdf, other]
Title: Guiding Large Language Models to Post-Edit Machine Translation with Error Annotations
Comments: 21 pages, 8 figures
Journal-ref: NAACL 2024 Findings
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI)
[193]  arXiv:2404.07840 [pdf, other]
Title: On Training Data Influence of GPT Models
Subjects: Computation and Language (cs.CL); Machine Learning (cs.LG)
[ total of 352 entries: 1-25 | ... | 94-118 | 119-143 | 144-168 | 169-193 | 194-218 | 219-243 | 244-268 | ... | 344-352 ]
[ showing 25 entries per page: fewer | more | all ]

Disable MathJax (What is MathJax?)

Links to: arXiv, form interface, find, cs, new, 2404, contact, help  (Access key information)