Artificial Intelligence

Papers filed under cs.AI on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,561 to 10,620 of 15,401

  1. Gating Before Commitment: Anticipating Intent Divergence to Prevent Post-Interaction Decision Failures in Autonomous Driving

    Cong Xu, Ravi Sankar

    cs.ROcs.AIarXiv:2608.26074v12026
  2. SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks

    Alexander Robey, Eric Wong, Hamed Hassani +1

    cs.LGcs.AIstat.MLarXiv:2310.03684v42023
  3. FRAME: separating sampling variation from representational cause in medical imaging fairness

    Mahshad Lotfinia, Daniel Truhn, Andreas Maier +1

    cs.CVcs.AIcs.LGarXiv:2608.25981v12026
  4. Finding and using interpretable latents in a neutrino foundation model with sparse autoencoders

    Raphaël Bonnet-Guerrini, Johann Ioannou-Nikolaides, Inar Timiryasov +1

    astro-ph.HEcs.AIcs.LGarXiv:2608.26090v12026
  5. Imitation Learning for Connection-Tableau Construction

    Fredrik Rømming, Mantas Bakšys, Martin S. Fixman +1

    cs.AIcs.LGcs.LOarXiv:2608.26009v12026
  6. SciMIF: Understanding Multimodal Instruction Following in Scientific Domains

    Ye Shen, Yuting Zheng, Dun Pei +4

    cs.AIcs.LGarXiv:2608.25973v12026
  7. How Robust Are Automated Fact-Checking Systems? A Cross-Benchmark Evaluation

    Aida Usmanova, Zangir Iklassov, Markus Leippold +1

    cs.AIcs.LGarXiv:2608.25934v12026
  8. Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

    Fuxiao Liu, Kevin Lin, Linjie Li +3

    cs.CVcs.AIcs.CEarXiv:2306.14565v42023
  9. Communication-Efficient Federated Deep Learning with Asynchronous Model Update and Temporally Weighted Aggregation

    Yang Chen, Xiaoyan Sun, Yaochu Jin

    cs.LGcs.AIcs.DCarXiv:1903.07424v12019
  10. AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining

    Haohe Liu, Yi Yuan, Xubo Liu +7

    cs.SDcs.AIcs.MMarXiv:2308.05734v32023
  11. MyoMechanix: Biomechanically-Grounded Compositional Skilled Activity Understanding and Coaching

    Hao Yin, Paritosh Parmar, Lijun Gu +6

    cs.CVcs.AIcs.ETarXiv:2608.26094v12026
  12. FuXi: A cascade machine learning forecasting system for 15-day global weather forecast

    Lei Chen, Xiaohui Zhong, Feng Zhang +4

    physics.ao-phcs.AIcs.LGarXiv:2306.12873v32023
  13. Difficulty-Aware Sample Allocation for Adaptive Data Augmentation in Semantic Segmentation

    Olasimbo Ayodeji Arigbabu, Abimbola Ismail Arigbabu

    cs.CVcs.AIarXiv:2608.25710v12026
  14. dm_control: Software and Tasks for Continuous Control

    Yuval Tassa, Saran Tunyasuvunakool, Alistair Muldal +8

    cs.ROcs.AIcs.LGarXiv:2006.12983v22020
  15. Formal Security Analysis of Neural Networks using Symbolic Intervals

    Shiqi Wang, Kexin Pei, Justin Whitehouse +2

    cs.AIcs.LOarXiv:1804.10829v32018
  16. Inductive Relation Prediction by Subgraph Reasoning

    Komal K. Teru, Etienne Denis, William L. Hamilton

    cs.LGcs.AIstat.MLarXiv:1911.06962v22019
  17. MeMark: Membrane-Space Watermarking for Spiking Neural Networks

    Roberto Riaño, Gorka Abad, Stjepan Picek +1

    cs.CRcs.AIcs.LGarXiv:2608.25738v12026
  18. Pointing the Way, Hiding the Destination: Practical Private Dense Retrieval at Scale

    Peichun Hua, Danyang Chen, Junan Zhang +5

    cs.CRcs.AIcs.IRarXiv:2608.25735v12026
  19. Unsupervised Anatomical Feature Learning via Diffusion Models: Enhanced Medical Image Segmentation with Denoising Diffusion Probabilistic Models

    Akshat G, Divyansh Gupta, Shaleen Bhatnagar +2

    cs.CVcs.AIcs.LGarXiv:2608.25693v12026
  20. DualOPSD: Adaptive Privileged Teachers for On-Policy Self-Distillation

    Yutong Chen, Guangfu Guo, Zhichao Xu +1

    cs.LGcs.AIarXiv:2608.26019v12026
  21. Deep Learning with Low Precision by Half-wave Gaussian Quantization

    Zhaowei Cai, Xiaodong He, Jian Sun +1

    cs.CVcs.AIcs.LGarXiv:1702.00953v12017
  22. Towards A Unified Information Bottleneck Framework for Time Series Explanations

    Xu Zheng, Zichuan Liu, Zhuomin Chen +7

    cs.LGcs.AIarXiv:2608.25897v12026
  23. TailSFT: Filtered Fine-Tuning Improves Post-Training Performance

    Sadhika Malladi, Samy Jelassi, Dylan Foster +2

    cs.LGcs.AIarXiv:2608.25756v12026
  24. VT-ADL: A Vision Transformer Network for Image Anomaly Detection and Localization

    Pankaj Mishra, Riccardo Verk, Daniele Fornasier +2

    cs.CVcs.AIcs.LGarXiv:2104.10036v12021
  25. Narcissus: Program Synthesis Using Context-Aware LLM Approximations

    Tilman Hinnerichs, Sebastijan Dumancic, Neil Yorke-Smith

    cs.AIcs.LGcs.PLarXiv:2608.25657v12026
  26. TraceML: An Empirical Analysis of Human-Agent Planning in Machine Learning Development

    Jiarui Yan, Weiwei Sun, Sijie Li +2

    cs.LGcs.AIarXiv:2608.26086v12026
  27. Measuring Faithfulness in Chain-of-Thought Reasoning

    Tamera Lanham, Anna Chen, Ansh Radhakrishnan +27

    cs.AIcs.CLcs.LGarXiv:2307.13702v12023
  28. Massively Parallel Methods for Deep Reinforcement Learning

    Arun Nair, Praveen Srinivasan, Sam Blackwell +11

    cs.LGcs.AIcs.DCarXiv:1507.04296v22015
  29. Towards Optimally Decentralized Multi-Robot Collision Avoidance via Deep Reinforcement Learning

    Pinxin Long, Tingxiang Fan, Xinyi Liao +3

    cs.ROcs.AIcs.LGarXiv:1709.10082v32017
  30. $R^3$: Training Robots to Reason in Natural Language via Reinforcement Learning

    Lehong Wu, Yuxiao Qu, Zheyuan Hu +4

    cs.ROcs.AIcs.CLarXiv:2608.26053v12026
  31. Git Re-Basin: Merging Models modulo Permutation Symmetries

    Samuel K. Ainsworth, Jonathan Hayase, Siddhartha Srinivasa

    cs.LGcs.AIarXiv:2209.04836v62022
  32. Trace Integrity for LLM Data Agents: A Vision for Auditable Structured Reasoning in Real-World Systems

    Srimonti Dutta, Akshata Kishore Moharir

    cs.AIcs.CLarXiv:2608.26036v12026
  33. It's a matter of timescale: non-linear utility in successor features and multi-objective planning and learning

    Liam P. H. Mertens, Lucas N. Alegre, Florent Delgrange +3

    cs.LGcs.AIarXiv:2608.25723v12026
  34. Sionna: An Open-Source Library for Next-Generation Physical Layer Research

    Jakob Hoydis, Sebastian Cammerer, Fayçal Ait Aoudia +4

    cs.ITcs.AIcs.LGarXiv:2203.11854v22022
  35. SwarmWorld: Stigmergic technological evolution in societies of language-model agents

    Subhadeep Pal, Fiona Y. Wang, Markus J. Buehler

    cs.AIcond-mat.mtrl-scics.CLarXiv:2608.26081v12026
  36. Learning Features by Watching Objects Move

    Deepak Pathak, Ross Girshick, Piotr Dollár +2

    cs.CVcs.AIcs.LGarXiv:1612.06370v22016
  37. How Much Rank Does LoRA Need? Rank-Error Bounds for Transformer Attention

    Gerard Conangla Planes

    cs.LGcs.AIcs.CLarXiv:2608.26052v12026
  38. On the Tractability of SHAP Explanations

    Guy Van den Broeck, Anton Lykov, Maximilian Schleich +1

    cs.AIcs.CCcs.LGarXiv:2009.08634v22020
  39. COMBO: Conservative Offline Model-Based Policy Optimization

    Tianhe Yu, Aviral Kumar, Rafael Rafailov +3

    cs.LGcs.AIcs.ROarXiv:2102.08363v22021
  40. Focal Self-attention for Local-Global Interactions in Vision Transformers

    Jianwei Yang, Chunyuan Li, Pengchuan Zhang +4

    cs.CVcs.AIcs.LGarXiv:2107.00641v12021
  41. One Symptom, Three Levers: A Critical Review of On-Policy Self-Distillation

    Justin Robert, Raheel Qader

    cs.LGcs.AIcs.CLarXiv:2608.25936v12026
  42. Inter-GPS: Interpretable Geometry Problem Solving with Formal Language and Symbolic Reasoning

    Pan Lu, Ran Gong, Shibiao Jiang +4

    cs.CLcs.AIcs.CVarXiv:2105.04165v32021
  43. AsymSpec: Context-Asymmetric Speculative Decoding for Agentic LLMs

    Sheng Liang, Yongyue Zhang, Nathanael Brian +4

    cs.AIcs.CLarXiv:2608.26004v12026
  44. Check Your Facts and Try Again: Improving Large Language Models with External Knowledge and Automated Feedback

    Baolin Peng, Michel Galley, Pengcheng He +8

    cs.CLcs.AIarXiv:2302.12813v32023
  45. Formal, Executable and Explainable Runtime Monitoring of Spoken Air Traffic Control Operational Procedures

    Roberto Luvini, Giacomo Longo, Alessandro Armando +1

    cs.AIcs.CLeess.ASarXiv:2608.25926v12026
  46. Prefix Sliding for efficient test-time scaling

    Niklas Muennighoff, Zhengyang Wang, Zeyi Chen +15

    cs.CLcs.AIcs.LGarXiv:2608.26070v12026
  47. ProcTHOR: Large-Scale Embodied AI Using Procedural Generation

    Matt Deitke, Eli VanderBilt, Alvaro Herrasti +8

    cs.AIcs.CVcs.ROarXiv:2206.06994v12022
  48. LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders

    Parishad BehnamGhader, Vaibhav Adlakha, Marius Mosbach +3

    cs.CLcs.AIarXiv:2404.05961v22024
  49. Structured Inference Networks for Nonlinear State Space Models

    Rahul G. Krishnan, Uri Shalit, David Sontag

    stat.MLcs.AIcs.LGarXiv:1609.09869v22016
  50. BERT: A Review of Applications in Natural Language Processing and Understanding

    M. V. Koroteev

    cs.CLcs.AIcs.LGarXiv:2103.11943v12021
  51. Visual Storytelling

    Ting-Hao, Huang, Francis Ferraro +13

    cs.CLcs.AIcs.CVarXiv:1604.03968v12016
  52. Query-Side Attacks on GNN-Based KGQA: Tracing Failures from Entity Linking to Answer Generation

    Pankaj Kumar, Subhankar Mishra

    cs.CLcs.AIcs.IRarXiv:2608.25922v12026
  53. Graph Retrieval-Augmented Generation: A Survey

    Boci Peng, Yun Zhu, Yongchao Liu +5

    cs.AIcs.CLcs.IRarXiv:2408.08921v22024
  54. Vote3Deep: Fast Object Detection in 3D Point Clouds Using Efficient Convolutional Neural Networks

    Martin Engelcke, Dushyant Rao, Dominic Zeng Wang +2

    cs.ROcs.AIcs.CVarXiv:1609.06666v22016
    Summaries:한국어
  55. VideoPoet: A Large Language Model for Zero-Shot Video Generation

    Dan Kondratyuk, Lijun Yu, Xiuye Gu +28

    cs.CVcs.AIarXiv:2312.14125v42023
  56. Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

    Bradly C. Stadie, Sergey Levine, Pieter Abbeel

    cs.AIcs.LGstat.MLarXiv:1507.00814v32015
  57. Reverse Curriculum Generation for Reinforcement Learning

    Carlos Florensa, David Held, Markus Wulfmeier +2

    cs.AIcs.LGcs.NEarXiv:1707.05300v32017
  58. Skill Issue: Are Skills Language-Invariant in LLMs?

    Bobby Cheng, Adam Gaber, Zhengyuan Liu +4

    cs.CLcs.AIcs.GTarXiv:2608.25832v12026
  59. BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains

    Yanis Labrak, Adrien Bazoge, Emmanuel Morin +3

    cs.CLcs.AIcs.LGarXiv:2402.10373v32024
  60. Unfolding Scientific Papers into Multi-Turn Generation Trajectories for Continued Pre-Training

    Qiankai Xu, Qiguang Chen, Zixin Su +4

    cs.CLcs.AIcs.LGarXiv:2608.25826v12026