Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

121 to 180 of 61,189

  1. iTool: Reinforced Fine-Tuning with Dynamic Deficiency Calibration for Advanced Tool Use

    Yirong Zeng, Xiao Ding, Yuxian Wang +8

    cs.CLcs.AIcs.LGarXiv:2501.09766v52025
  2. Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought

    Abulhair Saparov, He He

    cs.CLarXiv:2210.01240v42022
  3. Federated Learning on Non-IID Graphs via Structural Knowledge Sharing

    Yue Tan, Yixin Liu, Guodong Long +3

    cs.LGcs.AIcs.DCarXiv:2211.13009v12022
  4. Curriculum Temperature for Knowledge Distillation

    Zheng Li, Xiang Li, Lingfeng Yang +5

    cs.CVarXiv:2211.16231v32022
  5. Language Models as Agent Models

    Jacob Andreas

    cs.CLcs.MAarXiv:2212.01681v12022
    Summaries:한국어
  6. A 64-core mixed-signal in-memory compute chip based on phase-change memory for deep neural network inference

    Manuel Le Gallo, Riduan Khaddam-Aljameh, Milos Stanisavljevic +26

    cs.ETarXiv:2212.02872v12022
  7. One-Stage Cascade Refinement Networks for Infrared Small Target Detection

    Yimian Dai, Xiang Li, Fei Zhou +3

    cs.CVarXiv:2212.08472v22022
  8. Multi-modal Molecule Structure-text Model for Text-based Retrieval and Editing

    Shengchao Liu, Weili Nie, Chengpeng Wang +6

    cs.LGcs.CLq-bio.QMarXiv:2212.10789v32022
  9. Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training

    Song Lai, Haohan Zhao, Rong Feng +9

    cs.LGcs.AIcs.CLarXiv:2507.05386v62025
  10. Worst Case and Probabilistic Analysis of the 2-Opt Algorithm for the TSP

    Matthias Englert, Heiko Röglin, Berthold Vöcking

    cs.DSarXiv:2302.06889v12023
  11. Learning Performance-Improving Code Edits

    Alexander Shypula, Aman Madaan, Yimeng Zeng +7

    cs.SEcs.AIcs.LGarXiv:2302.07867v52023
  12. Can Pre-trained Vision and Language Models Answer Visual Information-Seeking Questions?

    Yang Chen, Hexiang Hu, Yi Luan +4

    cs.CVcs.AIcs.CLarXiv:2302.11713v52023
  13. Learning by Distilling Context

    Charlie Snell, Dan Klein, Ruiqi Zhong

    cs.CLcs.AIarXiv:2209.15189v12022
  14. LayoutDM: Discrete Diffusion Model for Controllable Layout Generation

    Naoto Inoue, Kotaro Kikuchi, Edgar Simo-Serra +2

    cs.CVcs.GRarXiv:2303.08137v12023
  15. GPT-4 Technical Report

    OpenAI, Josh Achiam, Steven Adler +278

    cs.CLcs.AIarXiv:2303.08774v62023
  16. FreeDoM: Training-Free Energy-Guided Conditional Diffusion Model

    Jiwen Yu, Yinhuai Wang, Chen Zhao +2

    cs.CVarXiv:2303.09833v12023
  17. LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

    Renrui Zhang, Jiaming Han, Chris Liu +7

    cs.CVcs.AIcs.CLarXiv:2303.16199v32023
  18. Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models

    Alejandro Lopez-Lira, Yuehua Tang

    q-fin.STcs.CLarXiv:2304.07619v62023
  19. Phase transition in Random Circuit Sampling

    A. Morvan, B. Villalonga, X. Mi +182

    quant-pharXiv:2304.11119v22023
  20. LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model

    Peng Gao, Jiaming Han, Renrui Zhang +9

    cs.CVcs.AIcs.CLarXiv:2304.15010v12023
  21. PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

    Xiaoman Zhang, Chaoyi Wu, Ziheng Zhao +4

    cs.CVarXiv:2305.10415v62023
  22. Enabling Large Language Models to Generate Text with Citations

    Tianyu Gao, Howard Yen, Jiatong Yu +1

    cs.CLcs.IRcs.LGarXiv:2305.14627v22023
  23. Towards Unified Text-based Person Retrieval: A Large-scale Multi-Attribute and Language Search Benchmark

    Shuyu Yang, Yinan Zhou, Yaxiong Wang +3

    cs.CVcs.MMarXiv:2306.02898v42023
  24. RS5M and GeoRSCLIP: A Large Scale Vision-Language Dataset and A Large Vision-Language Model for Remote Sensing

    Zilun Zhang, Tiancheng Zhao, Yulong Guo +1

    cs.CVcs.AIcs.CLarXiv:2306.11300v52023
  25. Matching Patients to Clinical Trials with Large Language Models

    Qiao Jin, Zifeng Wang, Charalampos S. Floudas +7

    cs.CLcs.AIarXiv:2307.15051v52023
  26. Studying Large Language Model Generalization with Influence Functions

    Roger Grosse, Juhan Bae, Cem Anil +14

    cs.LGcs.CLstat.MLarXiv:2308.03296v12023
  27. GeoCLIP: Clip-Inspired Alignment between Locations and Images for Effective Worldwide Geo-localization

    Vicente Vivanco Cepeda, Gaurav Kumar Nayak, Mubarak Shah

    cs.CVcs.LGarXiv:2309.16020v22023
  28. Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

    Yangsibo Huang, Samyak Gupta, Mengzhou Xia +2

    cs.CLcs.AIcs.CRarXiv:2310.06987v12023
  29. CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving

    Yuhan Liu, Hanchen Li, Yihua Cheng +11

    cs.NIcs.LGarXiv:2310.07240v62023
  30. Reaching the Limit in Autonomous Racing: Optimal Control versus Reinforcement Learning

    Yunlong Song, Angel Romero, Matthias Mueller +2

    cs.ROcs.LGarXiv:2310.10943v22023
  31. Linear Representations of Sentiment in Large Language Models

    Curt Tigges, Oskar John Hollinsworth, Atticus Geiger +1

    cs.LGcs.AIcs.CLarXiv:2310.15154v12023
  32. DeepInception: Hypnotize Large Language Model to Be Jailbreaker

    Xuan Li, Zhanke Zhou, Jianing Zhu +3

    cs.LGcs.CRarXiv:2311.03191v52023
  33. Neural General Circulation Models for Weather and Climate

    Dmitrii Kochkov, Janni Yuval, Ian Langmore +13

    physics.ao-phcs.LGphysics.comp-pharXiv:2311.07222v32023
  34. Gibbs Sampling with People

    Peter M. C. Harrison, Raja Marjieh, Federico Adolfi +5

    q-bio.NCcs.AIcs.CVarXiv:2008.02595v22020
  35. Halo Formation in Warm Dark Matter Models

    Paul Bode, Jeremiah P. Ostriker, Neil Turok

    astro-pharXiv:astro-ph/0010389v32000
  36. Large Language Models for Robotics: A Survey

    Fanlong Zeng, Wensheng Gan, Zezheng Huai +5

    cs.ROcs.AIarXiv:2311.07226v22023
  37. Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning

    Rohit Girdhar, Mannat Singh, Andrew Brown +7

    cs.CVcs.AIcs.GRarXiv:2311.10709v22023
  38. HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data

    Qifan Yu, Juncheng Li, Longhui Wei +6

    cs.CVcs.AIarXiv:2311.13614v22023
  39. VBench: Comprehensive Benchmark Suite for Video Generative Models

    Ziqi Huang, Yinan He, Jiashuo Yu +13

    cs.CVarXiv:2311.17982v12023
  40. The AI Assessment Scale (AIAS): A Framework for Ethical Integration of Generative AI in Educational Assessment

    Mike Perkins, Leon Furze, Jasper Roe +1

    cs.AIarXiv:2312.07086v22023
  41. Large Legal Fictions: Profiling Legal Hallucinations in Large Language Models

    Matthew Dahl, Varun Magesh, Mirac Suzgun +1

    cs.CLcs.AIcs.CYarXiv:2401.01301v22024
  42. FlightLLM: Efficient Large Language Model Inference with a Complete Mapping Flow on FPGAs

    Shulin Zeng, Jun Liu, Guohao Dai +14

    cs.ARcs.AIarXiv:2401.03868v22024
  43. One Polluted Page Is Enough: Evaluating Web Content Pollution in LLM Recommenders

    Minghao Luo, Liang Chen

    cs.CLcs.AIarXiv:2606.13610v22026
  44. i1: A Simple and Fully Open Recipe for Strong Text-to-Image Models

    Boya Zeng, Tianze Luo, Shu Pu +4

    cs.CVarXiv:2606.11289v12026
  45. SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories

    Zhuoyun Yu, Xin Xie, Wuguannan Yao +4

    cs.CLcs.AIcs.LGarXiv:2606.01311v12026
  46. Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

    Josef Chen

    cs.ARcs.AIcs.DCarXiv:2605.30571v12026
  47. The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages

    Eric Onyame, Runtao Zhou, Kowshik Thopalli +2

    cs.CLcs.AIarXiv:2605.27901v12026
  48. The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence

    Aili Chen, Aonian Li, Baichuan Zhou +215

    cs.AIcs.CLcs.LGarXiv:2605.26494v22026
  49. Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization

    Anmol Agarwal, Natalie Neamtu, Pranjal Aggarwal +6

    cs.SEcs.AIcs.CLarXiv:2605.26457v12026
  50. Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

    Zachary Novack, Stephen Brade, Haven Kim +8

    cs.SDcs.AIcs.LGarXiv:2605.22717v12026
  51. Forecasting Downstream Performance of LLMs With Proxy Metrics

    Arkil Patel, Siva Reddy, Marius Mosbach +1

    cs.CLcs.LGarXiv:2605.18607v12026
  52. OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization

    Zhongzhu Zhou, Donglin Zhuang, Jisen Li +4

    cs.LGcs.AIcs.DCarXiv:2605.17757v12026
  53. ALAS: Transactional and Dynamic Multi-Agent LLM Planning

    Longling Geng, Edward Y. Chang

    cs.MAarXiv:2511.03094v12025
  54. A Categorical Quantum Logic

    Samson Abramsky, Ross Duncan

    quant-phcs.LOarXiv:quant-ph/0512114v12005
  55. Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?

    Yichen Feng, Yuetai Li, Chunjiang Liu +14

    cs.CVcs.AIcs.HCarXiv:2605.12684v12026
  56. Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning

    Junhao Shen, Teng Zhang, Xiaoyan Zhao +1

    cs.LGcs.CLarXiv:2605.10923v22026
  57. MemPrivacy: Privacy-Preserving Personalized Memory Management for Edge-Cloud Agents

    Yining Chen, Jihao Zhao, Bo Tang +5

    cs.CRcs.CLarXiv:2605.09530v32026
  58. AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents

    Nithin Somasekharan, Rabi Pathak, Manushri Dhanakoti +4

    physics.flu-dyncs.AIarXiv:2605.06607v32026
  59. Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe

    Wenjin Hou, Shangpin Peng, Weinong Wang +13

    cs.LGarXiv:2605.03677v22026
  60. A Survey on Metaverse: Fundamentals, Security, and Privacy

    Yuntao Wang, Zhou Su, Ning Zhang +4

    cs.CRarXiv:2203.02662v42022