Artificial Intelligence
Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,561 to 4,620 of 15,246
Atom of Thoughts for Markov LLM Test-Time Scaling
Fengwei Teng, Quan Shi, Zhaoyang Yu +4
cs.CLcs.AIcs.LGarXiv:2502.12018v42025MAPP: a Scalable Multi-Agent Path Planning Algorithm with Tractability and Completeness Guarantees
Ko-Hsin Cindy Wang, Adi Botea
cs.AIarXiv:1401.3905v12014Geometry-aware Latent Autoregressive Generative Model for PDEs in Complex Domains
Zi Wang, Minghui Xu, Tapan Mukerji
cs.LGcs.AIarXiv:2609.00297v12026CoLT-Drive: Counterfactual Long-Tail Benchmarking and Knowledge-Preserving Adaptation for Driving Affordance Prediction
Zhengxu Tang, Guofeng Cui, Ziyu Gong +8
cs.CVcs.AIcs.CLarXiv:2609.00242v12026UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent
Jianke Zhang, Yanjiang Guo, Yucheng Hu +3
cs.CVcs.AIarXiv:2501.18867v32025Maximizing Confidence Alone Improves Reasoning
Mihir Prabhudesai, Lili Chen, Alex Ippoliti +3
cs.LGcs.AIarXiv:2505.22660v42025Neural probabilistic motor primitives for humanoid control
Josh Merel, Leonard Hasenclever, Alexandre Galashov +5
cs.LGcs.AIcs.ROarXiv:1811.11711v22018Quantization Hurts Reasoning? An Empirical Study on Quantized Reasoning Models
Ruikang Liu, Yuxuan Sun, Manyi Zhang +5
cs.CLcs.AIarXiv:2504.04823v22025Learning Neural Causal Models from Unknown Interventions
Nan Rosemary Ke, Olexa Bilaniuk, Anirudh Goyal +6
stat.MLcs.AIcs.LGarXiv:1910.01075v22019How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks
Longju Bai, Zhemin Huang, Xingyao Wang +5
cs.CLcs.AIcs.CYarXiv:2604.22750v22026Degradation-Aware Feature Perturbation for All-in-One Image Restoration
Xiangpeng Tian, Xiangyu Liao, Xiao Liu +2
cs.CVcs.AIarXiv:2505.12630v12025Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning
Shulin Tian, Ruiqi Wang, Hongming Guo +7
cs.CVcs.AIarXiv:2506.13654v12025Hard Sample Aware Network for Contrastive Deep Graph Clustering
Yue Liu, Xihong Yang, Sihang Zhou +7
cs.LGcs.AIarXiv:2212.08665v32022TeMP: Temporal Message Passing for Temporal Knowledge Graph Completion
Jiapeng Wu, Meng Cao, Jackie Chi Kit Cheung +1
cs.LGcs.AIcs.CLarXiv:2010.03526v12020Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning
Xinbin Yuan, Jian Zhang, Kaixin Li +8
cs.AIarXiv:2505.12370v22025Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?
Zhiyuan Zeng, Qinyuan Cheng, Zhangyue Yin +2
cs.LGcs.AIcs.CLarXiv:2502.12215v22025A General Non-Probabilistic Theory of Inductive Reasoning
Wolfgang Spohn
cs.AIarXiv:1304.2375v12013Dynamic Dual-Granularity Skill Bank for Agentic RL
Songjun Tu, Chengdong Xu, Qichao Zhang +5
cs.AIarXiv:2603.28716v22026Commonsense for Generative Multi-Hop Question Answering Tasks
Lisa Bauer, Yicheng Wang, Mohit Bansal
cs.CLcs.AIarXiv:1809.06309v32018AI Research Agents for Machine Learning: Search, Exploration, and Generalization in MLE-bench
Edan Toledo, Karen Hambardzumyan, Martin Josifoski +22
cs.AIcs.LGarXiv:2507.02554v22025Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
Ekaterina Fadeeva, Aleksandr Rubashevskii, Artem Shelmanov +9
cs.CLcs.AIcs.LGarXiv:2403.04696v22024ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents
Qiuchen Wang, Ruixue Ding, Zehui Chen +4
cs.CVcs.AIcs.CLarXiv:2502.18017v22025Channel-Wise Attention-Based Network for Self-Supervised Monocular Depth Estimation
Jiaxing Yan, Hong Zhao, Penghui Bu +1
cs.CVcs.AIcs.LGarXiv:2112.13047v12021BranchGRPO: Stable and Efficient GRPO with Structured Branching in Diffusion Models
Yuming Li, Yikai Wang, Yuying Zhu +4
cs.CVcs.AIcs.LGarXiv:2509.06040v52025Between Underthinking and Overthinking: An Empirical Study of Reasoning Length and correctness in LLMs
Jinyan Su, Jennifer Healey, Preslav Nakov +1
cs.CLcs.AIarXiv:2505.00127v12025ShorterBetter: Guiding Reasoning Models to Find Optimal Inference Length for Efficient Reasoning
Jingyang Yi, Jiazheng Wang, Sida Li
cs.AIarXiv:2504.21370v42025NFormer: Robust Person Re-identification with Neighbor Transformer
Haochen Wang, Jiayi Shen, Yongtuo Liu +2
cs.CVcs.AIarXiv:2204.09331v12022Bayesian Network Constraint-Based Structure Learning Algorithms: Parallel and Optimised Implementations in the bnlearn R Package
Marco Scutari
stat.COcs.AIcs.MSarXiv:1406.7648v22014Improving the Diffusability of Autoencoders
Ivan Skorokhodov, Sharath Girish, Benran Hu +5
cs.CVcs.AIcs.LGarXiv:2502.14831v32025MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers
Chaithanya Bandi, Razvan-Gabriel Dumitru, Ben Hertzberg +20
cs.SEcs.AIarXiv:2602.00933v32026A Deep-Reinforcement Learning Approach for Software-Defined Networking Routing Optimization
Giorgio Stampa, Marta Arias, David Sanchez-Charles +2
cs.NIcs.AIarXiv:1709.07080v12017Adjoint Sampling: Highly Scalable Diffusion Samplers via Adjoint Matching
Aaron Havens, Benjamin Kurt Miller, Bing Yan +10
cs.LGcs.AIarXiv:2504.11713v32025"Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills in the Wild
Yi Liu, Zhihao Chen, Yanjun Zhang +4
cs.CRcs.AIcs.CLarXiv:2602.06547v42026SAMGPT: Text-free Graph Foundation Model for Multi-domain Pre-training and Cross-domain Adaptation
Xingtong Yu, Zechuan Gong, Chang Zhou +2
cs.CLcs.AIarXiv:2502.05424v22025Edge-Cloud Collaborative Computing on Distributed Intelligence and Model Optimization: A Survey
Jing Liu, Yao Du, Kun Yang +8
cs.DCcs.AIcs.LGarXiv:2505.01821v52025Don't be lazy: CompleteP enables compute-efficient deep transformers
Nolan Dey, Bin Claire Zhang, Lorenzo Noci +6
cs.LGcs.AIarXiv:2505.01618v42025SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning
Rui Pan, Yinwei Dai, Zhihao Zhang +3
cs.LGcs.AIarXiv:2504.07891v22025DDT: Decoupled Diffusion Transformer
Shuai Wang, Zhi Tian, Weilin Huang +1
cs.CVcs.AIarXiv:2504.05741v22025SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience
Zeyi Sun, Ziyu Liu, Yuhang Zang +5
cs.AIcs.CLcs.CVarXiv:2508.04700v22025Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness
Felix Friedrich, Manuel Brack, Lukas Struppek +4
cs.LGcs.AIcs.CVarXiv:2302.10893v32023JSONSchemaBench: A Rigorous Benchmark of Structured Outputs for Language Models
Saibo Geng, Hudson Cooper, Michał Moskal +6
cs.CLcs.AIarXiv:2501.10868v32025Exploiting Sample Uncertainty for Domain Adaptive Person Re-Identification
Kecheng Zheng, Cuiling Lan, Wenjun Zeng +2
cs.CVcs.AIarXiv:2012.08733v22020Radial Attention: $O(n\log n)$ Sparse Attention with Energy Decay for Long Video Generation
Xingyang Li, Muyang Li, Tianle Cai +11
cs.CVcs.AIcs.LGarXiv:2506.19852v22025Rethinking Learnability in Offline Data-driven Optimization
Chao Qian, Chen-Guang Wang, Rong-Xi Tan +1
cs.LGcs.AIcs.NEarXiv:2609.01493v22026Fully Parameterized Quantile Function for Distributional Reinforcement Learning
Derek Yang, Li Zhao, Zichuan Lin +3
cs.LGcs.AIstat.MLarXiv:1911.02140v32019A survey of algorithmic recourse: definitions, formulations, solutions, and prospects
Amir-Hossein Karimi, Gilles Barthe, Bernhard Schölkopf +1
cs.LGcs.AIstat.MLarXiv:2010.04050v22020ReasonIR: Training Retrievers for Reasoning Tasks
Rulin Shao, Rui Qiao, Varsha Kishore +8
cs.AIcs.CLcs.IRarXiv:2504.20595v12025HiveTraceGuard-Pro: A Compact Generative Guardrail for Prompt Injection, Jailbreaks, and Adversarial Obfuscation
Nikita Oblakov, Sabrina Sadiekh, Evgeniy Kokuykin
cs.CRcs.AIarXiv:2609.01046v12026TxGemma: Efficient and Agentic LLMs for Therapeutics
Eric Wang, Samuel Schmidgall, Paul F. Jaeger +6
cs.AIcs.CLcs.LGarXiv:2504.06196v12025Interactive Debugging and Steering of Multi-Agent AI Systems
Will Epperson, Gagan Bansal, Victor Dibia +4
cs.MAcs.AIcs.HCarXiv:2503.02068v12025Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
Hongjin Su, Ruoxi Sun, Jinsung Yoon +3
cs.LGcs.AIarXiv:2501.10893v12025VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos
Xubin Ren, Lingrui Xu, Long Xia +3
cs.IRcs.AIcs.CVarXiv:2502.01549v12025Who Judges the Judges? A Chinese Safety QA Benchmark for Evaluating LLM Responses and Safety Judges
Rui Yang, Shuang Huang, Junhua Liu +7
cs.CRcs.AIarXiv:2609.01210v12026Adversarial Laser Beam: Effective Physical-World Attack to DNNs in a Blink
Ranjie Duan, Xiaofeng Mao, A. K. Qin +4
cs.LGcs.AIcs.CRarXiv:2103.06504v12021Self-Training Elicits Concise Reasoning in Large Language Models
Tergel Munkhbat, Namgyu Ho, Seo Hyun Kim +3
cs.CLcs.AIcs.LGarXiv:2502.20122v32025Causal Evidentiary Governance for High-Risk Machine Learning Systems
Samah Kareem, Barış Çeliktaş
cs.CYcs.AIarXiv:2609.01040v12026AI-Researcher: Autonomous Scientific Innovation
Jiabin Tang, Lianghao Xia, Zhonghang Li +1
cs.AIarXiv:2505.18705v12025Vision-Language-Guided Pseudo-Labels for Unsupervised Domain Adaptation in Semantic Segmentation for Waste Sorting
Udo Schlegel, Shubhangi, Gabriel Dax +3
cs.CVcs.AIcs.LGarXiv:2609.00898v12026EvoSCM: Scientific Belief Revision Through Causal Model Evolution and Experimentation
Qing Zhao, Haowei Li, Weijian Deng +2
cs.AIarXiv:2609.01526v12026FlowDPS: Flow-Driven Posterior Sampling for Inverse Problems
Jeongsol Kim, Bryan Sangwoo Kim, Jong Chul Ye
cs.CVcs.AIcs.LGarXiv:2503.08136v12025