Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
241 to 300 of 15,222
AgentInstruct: Toward Generative Teaching with Agentic Flows
Arindam Mitra, Luciano Del Corro, Guoqing Zheng +11
cs.AIcs.CLcs.LGarXiv:2407.03502v12024Substructure Discovery Using Minimum Description Length and Background Knowledge
D. J. Cook, L. B. Holder
cs.AIarXiv:cs/9402102v11994BoardgameQA: A Dataset for Natural Language Reasoning with Contradictory Information
Mehran Kazemi, Quan Yuan, Deepti Bhatia +4
cs.CLcs.AIcs.LGarXiv:2306.07934v12023Soft Tokens, Hard Truths
Natasha Butt, Ariel Kwiatkowski, Ismail Labiad +2
cs.CLcs.AIcs.LGarXiv:2509.19170v22025Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon Captioning
Jifan Zhang, Lalit Jain, Yang Guo +9
cs.LGcs.AIcs.CLarXiv:2406.10522v22024Backward Lens: Projecting Language Model Gradients into the Vocabulary Space
Shahar Katz, Yonatan Belinkov, Mor Geva +1
cs.CLcs.AIcs.LGarXiv:2402.12865v12024Beyond English-Centric LLMs: What Language Do Multilingual Language Models Think in?
Chengzhi Zhong, Fei Cheng, Qianying Liu +5
cs.CLcs.AIarXiv:2408.10811v12024Addition in Four Movements: Mapping Layer-wise Information Trajectories in LLMs
Yao Yan
cs.AIarXiv:2506.07824v22025Learning with Challenges: Adaptive Difficulty-Aware Data Generation for Mobile GUI Agent Training
Linjia Kang, Zhimin Wang, Yongkang Zhang +5
cs.AIarXiv:2601.22781v12026MI9: An Integrated Runtime Governance Framework for Agentic AI
Charles L. Wang, Trisha Singhal, Ameya Kelkar +1
cs.AIcs.ETcs.MAarXiv:2508.03858v42025New Bounds for Zarankiewicz Numbers via Reinforced LLM Evolutionary Search
Jay Bhan, Nicole Nobili, Patrick Langer
cs.AImath.COarXiv:2605.01120v22026TinyV: Reducing False Negatives in Verification Improves RL for LLM Reasoning
Zhangchen Xu, Yuetai Li, Fengqing Jiang +4
cs.LGcs.AIcs.CLarXiv:2505.14625v22025Brain-to-Text Decoding: A Non-invasive Approach via Typing
Jarod Lévy, Mingfang Zhang, Svetlana Pinet +4
eess.SPcs.AIcs.CLarXiv:2502.17480v12025Multi-Task Federated Reinforcement Learning with Adversaries
Aqeel Anwar, Arijit Raychowdhury
cs.LGcs.AIarXiv:2103.06473v12021SemGloVe: Semantic Co-occurrences for GloVe from BERT
Leilei Gan, Zhiyang Teng, Yue Zhang +3
cs.CLcs.AIarXiv:2012.15197v22020Predictable Scale: Part I, Step Law -- Optimal Hyperparameter Scaling Law in Large Language Model Pretraining
Houyi Li, Wenzhen Zheng, Qiufeng Wang +10
cs.LGcs.AIarXiv:2503.04715v72025Learn Before Represent: Bridging Generative and Contrastive Learning for Domain-Specific LLM Embeddings
Xiaoyu Liang, Yuchen Peng, Jiale Luo +3
cs.IRcs.AIarXiv:2601.11124v12026Property Prediction of Stacked Bilayer Materials: A Multimodal Learning Approach
An Vuong, Minh-Hao Van, Chen Zhao +1
cs.AIcond-mat.mtrl-sciarXiv:2606.01012v12026MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions
Kai Zhang, Yi Luan, Hexiang Hu +5
cs.CVcs.AIcs.CLarXiv:2403.19651v22024Reasoning Up the Instruction Ladder for Controllable Language Models
Zishuo Zheng, Vidhisha Balachandran, Chan Young Park +2
cs.CLcs.AIarXiv:2511.04694v52025CAFe: Unifying Representation and Generation with Contrastive-Autoregressive Finetuning
Hao Yu, Zhuokai Zhao, Shen Yan +7
cs.CVcs.AIcs.CLarXiv:2503.19900v12025Proof-or-Stop: Don't Trust the Agent, Trust the Evidence -- Loop Engineering for Verifiable Evidence-Gated Lifecycle Control
Jek Huang, Jeffery Hsia, Jiayi Sun +3
cs.AIcs.SEarXiv:2607.14890v12026Graph-Augmented Large Language Model Agents: Current Progress and Future Prospects
Yixin Liu, Guibin Zhang, Kun Wang +2
cs.AIarXiv:2507.21407v22025q-Learning in Continuous Time
Yanwei Jia, Xun Yu Zhou
cs.LGcs.AIq-fin.CParXiv:2207.00713v42022STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems
Akash Bonagiri, Gerard Janno Anderias, Saee Patil +6
cs.LGcs.AIarXiv:2605.02122v22026Don't Complete It! Preventing Unhelpful Code Completion for Productive and Sustainable Neural Code Completion Systems
Zhensu Sun, Xiaoning Du, Fu Song +4
cs.SEcs.AIarXiv:2209.05948v32022RepMLPNet: Hierarchical Vision MLP with Re-parameterized Locality
Xiaohan Ding, Honghao Chen, Xiangyu Zhang +2
cs.CVcs.AIcs.LGarXiv:2112.11081v22021Progressive Distillation for Fast Sampling of Diffusion Models
Tim Salimans, Jonathan Ho
cs.LGcs.AIstat.MLarXiv:2202.00512v22022Learning to Reason for Hallucination Span Detection
Hsuan Su, Ting-Yao Hu, Hema Swetha Koppula +7
cs.CLcs.AIcs.LGarXiv:2510.02173v22025Continual Contrastive Learning for Image Classification
Zhiwei Lin, Yongtao Wang, Hongxiang Lin
cs.CVcs.AIarXiv:2107.01776v42021LPASS: Linear Probes as Stepping Stones for vulnerability detection using compressed LLMs
Luis Ibanez-Lissen, Lorena Gonzalez-Manzano, Jose Maria de Fuentes +1
cs.CRcs.AIarXiv:2505.24451v12025TikZilla: Scaling Text-to-TikZ with High-Quality Data and Reinforcement Learning
Christian Greisinger, Steffen Eger
cs.AIcs.CLcs.CVarXiv:2603.03072v32026DiagrammerGPT: Generating Open-Domain, Open-Platform Diagrams via LLM Planning
Abhay Zala, Han Lin, Jaemin Cho +1
cs.CVcs.AIcs.CLarXiv:2310.12128v22023Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents
Utsav Garg, Sungjin Hong, Jason Jung +6
cs.AIcs.LGarXiv:2606.31648v12026The Probabilities Also Matter: A More Faithful Metric for Faithfulness of Free-Text Explanations in Large Language Models
Noah Y. Siegel, Oana-Maria Camburu, Nicolas Heess +1
cs.CLcs.AIarXiv:2404.03189v22024Reaching Human-level Performance in Automatic Grammatical Error Correction: An Empirical Study
Tao Ge, Furu Wei, Ming Zhou
cs.CLcs.AIarXiv:1807.01270v52018Improved lower bounds for the Shannon capacity of odd cycles
Nathaniel Itty, Christopher D. Rosin, Chase Carstensen +1
cs.ITcs.AIcs.DMarXiv:2607.21517v22026Fair Adversarial Gradient Tree Boosting
Vincent Grari, Boris Ruf, Sylvain Lamprier +1
cs.LGcs.AIcs.CYarXiv:1911.05369v22019Estimating individual treatment effect: generalization bounds and algorithms
Uri Shalit, Fredrik D. Johansson, David Sontag
stat.MLcs.AIcs.LGarXiv:1606.03976v52016Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers
Xin-Qiang Cai, Wei Wang, Feng Liu +3
cs.LGcs.AIarXiv:2510.00915v42025Scratchy: Visual-Scratchpad Multimodal Reasoning for Cryptographic Proof Generation in EasyCrypt
Yupeng Ren, Zhaoxuan Li, Rui Zhang
cs.CRcs.AIarXiv:2609.06226v12026OceanLight: Efficient Global Ocean Forecasting via Geometry-Adaptive Unstructured Mesh Representation
Wei Wu, Xiang Wang, Hongze Leng +3
cs.LGcs.AIarXiv:2608.16070v12026Paper2Agent: Reimagining Research Papers As Interactive and Reliable AI Agents
Jiacheng Miao, Joe R. Davis, Yaohui Zhang +2
cs.AIcs.CLcs.LGarXiv:2509.06917v22025Cross-Modal Causal Relational Reasoning for Event-Level Visual Question Answering
Yang Liu, Guanbin Li, Liang Lin
cs.CVcs.AIarXiv:2207.12647v82022Multimodal Whole Slide Foundation Model for Pathology
Tong Ding, Sophia J. Wagner, Andrew H. Song +20
eess.IVcs.AIcs.CVarXiv:2411.19666v12024Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training
Evan Hubinger, Carson Denison, Jesse Mu +36
cs.CRcs.AIcs.CLarXiv:2401.05566v32024DNABERT-2: Efficient Foundation Model and Benchmark For Multi-Species Genome
Zhihan Zhou, Yanrong Ji, Weijian Li +3
q-bio.GNcs.AIcs.CEarXiv:2306.15006v22023Wav2Letter: an End-to-End ConvNet-based Speech Recognition System
Ronan Collobert, Christian Puhrsch, Gabriel Synnaeve
cs.LGcs.AIcs.CLarXiv:1609.03193v22016Subject-driven Text-to-Image Generation via Apprenticeship Learning
Wenhu Chen, Hexiang Hu, Yandong Li +4
cs.CVcs.AIarXiv:2304.00186v52023Graph of Thoughts: Solving Elaborate Problems with Large Language Models
Maciej Besta, Nils Blach, Ales Kubicek +8
cs.CLcs.AIcs.LGarXiv:2308.09687v42023The Impact of Positional Encoding on Length Generalization in Transformers
Amirhossein Kazemnejad, Inkit Padhi, Karthikeyan Natesan Ramamurthy +2
cs.CLcs.AIcs.LGarXiv:2305.19466v22023Beyond neural scaling laws: beating power law scaling via data pruning
Ben Sorscher, Robert Geirhos, Shashank Shekhar +2
cs.LGcs.AIcs.CVarXiv:2206.14486v62022Same Request, Different Boundary: Evaluating Cybersecurity Assistance across Conversational Contexts
Rui Yang, Yang Hong, Yichao Xu +3
cs.AIcs.CRarXiv:2609.00578v12026Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
A. Sophia Koepke, Daniil Zverev, Shiry Ginosar +1
cs.CVcs.AIcs.LGarXiv:2604.18572v22026Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers
Harsh Kohli, Srinivasan Parthasarathy, Huan Sun +1
cs.CLcs.AIcs.LGarXiv:2604.07822v22026PLDR-LLMs Reason At Self-Organized Criticality
Burc Gokden
cs.AIcs.CLcs.LGarXiv:2603.23539v12026Deep Learning-Based Multi-User Communication Design for Dense IoT Networks: Interference-Aware Finite-Blocklength Communication and Preliminary MIMO Extensions
Arkadeep Sinha, Shubham Paul, R. Manivasakan
cs.ITcs.AIarXiv:2608.22923v12026Participatory Moral AI Is Not Neutral: The Invisible Hand of Developers
Taenyun Kim, Edyta Bogucka, Daniele Quercia
cs.AIarXiv:2608.14522v12026Reinforcement Learning for Code Optimization
Pierre Chambon, Kunhao Zheng, Juliette Decugis +2
cs.LGcs.AIarXiv:2607.25970v12026CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning
Eric Onyame, Akash Ghosh, Subhadip Baidya +3
cs.AIcs.CLarXiv:2601.13262v22026