Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

22,261 to 22,320 of 61,134

  1. Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions

    Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

    cs.CVcs.CLarXiv:1811.10652v32018
  2. No Language Left Behind: Scaling Human-Centered Machine Translation

    NLLB Team, Marta R. Costa-jussà, James Cross +36

    cs.CLcs.AIarXiv:2207.04672v32022
  3. LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization

    Xueyang Zhou, Yangming Xu, Guiyao Tie +5

    cs.CVcs.ROarXiv:2510.03827v22025
  4. Spatial-Angular Interaction for Light Field Image Super-Resolution

    Yingqian Wang, Longguang Wang, Jungang Yang +3

    eess.IVcs.CVarXiv:1912.07849v32019
  5. Simple Contrastive Graph Clustering

    Yue Liu, Xihong Yang, Sihang Zhou +1

    cs.LGcs.AIarXiv:2205.07865v32022
  6. Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective

    Kun Chen, Peng Shi, Fanfan Liu +4

    cs.LGcs.AIcs.CLarXiv:2602.09782v22026
  7. Trust The Typical

    Debargha Ganguly, Sreehari Sankar, Biyao Zhang +8

    cs.CLcs.AIcs.DCarXiv:2602.04581v12026
  8. Does "AI" stand for augmenting inequality in the era of covid-19 healthcare?

    David Leslie, Anjali Mazumder, Aidan Peppin +2

    cs.CYcs.LGarXiv:2105.07844v12021
  9. Large Language Model Agent: A Survey on Methodology, Applications and Challenges

    Junyu Luo, Weizhi Zhang, Ye Yuan +23

    cs.CLarXiv:2503.21460v12025
  10. Agentic AI: A Comprehensive Survey of Architectures, Applications, and Future Directions

    Mohamad Abou Ali, Fadi Dornaika

    cs.AIcs.LGarXiv:2510.25445v12025
  11. Context Learning for Multi-Agent Discussion

    Xingyuan Hua, Sheng Yue, Xinyi Li +3

    cs.AIcs.LGcs.MAarXiv:2602.02350v32026
  12. Steering LLMs via Scalable Interactive Oversight

    Enyu Zhou, Zhiheng Xi, Long Ma +9

    cs.AIcs.LGarXiv:2602.04210v22026
  13. Sensor Fault Detection, Isolation and Identification Using Multiple Model-based Hybrid Kalman Filter for Gas Turbine Engines

    Bahareh Pourbabaee, Nader Meskin, Khashayar Khorasani

    eess.SYarXiv:1505.02063v22015
  14. Alternating Reinforcement Learning for Rubric-Based Reward Modeling in Non-Verifiable LLM Post-Training

    Ran Xu, Tianci Liu, Zihan Dong +6

    cs.CLcs.LGarXiv:2602.01511v22026
  15. Multi-agent Architecture Search via Agentic Supernet

    Guibin Zhang, Luyang Niu, Junfeng Fang +3

    cs.LGcs.CLcs.MAarXiv:2502.04180v22025
  16. FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale

    Ajay Patel, Colin Raffel, Chris Callison-Burch

    cs.CLcs.LGarXiv:2601.22146v32026
  17. Ethics-Based Auditing to Develop Trustworthy AI

    Jakob Mokander, Luciano Floridi

    cs.CYcs.AIarXiv:2105.00002v12021
  18. Deep Learning Object Detection Methods for Ecological Camera Trap Data

    Stefan Schneider, Graham W. Taylor, Stefan C. Kremer

    cs.CVarXiv:1803.10842v12018
  19. VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agents

    Zirui Wang, Junyi Zhang, Jiaxin Ge +9

    cs.CVarXiv:2601.16973v12026
  20. RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete

    Yuheng Ji, Huajie Tan, Jiayu Shi +14

    cs.ROcs.CVarXiv:2502.21257v22025
  21. Your Brain on ChatGPT: Accumulation of Cognitive Debt when Using an AI Assistant for Essay Writing Task

    Nataliya Kosmyna, Eugene Hauptmann, Ye Tong Yuan +5

    cs.AIarXiv:2506.08872v22025
  22. InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning

    Matthew Y. R. Yang, Hao Bai, Ian Wu +3

    cs.LGcs.AIcs.CLarXiv:2601.14209v12026
  23. Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

    Vidya Srinivas, Zachary Englhardt, Vikram Iyer +1

    cs.CLarXiv:2511.07397v32025
    Summaries:한국어
  24. A Generative Appearance Model for End-to-end Video Object Segmentation

    Joakim Johnander, Martin Danelljan, Emil Brissman +2

    cs.CVarXiv:1811.11611v22018
  25. A Survey on LLM-as-a-Judge

    Jiawei Gu, Xuhui Jiang, Zhichao Shi +13

    cs.CLcs.AIarXiv:2411.15594v62024
  26. An All-in-One Network for Dehazing and Beyond

    Boyi Li, Xiulian Peng, Zhangyang Wang +2

    cs.CVcs.AIarXiv:1707.06543v12017
  27. Introduction to Machine Learning

    Laurent Younes

    stat.MLcs.LGarXiv:2409.02668v22024
  28. Flexible-Antenna Systems: A Pinching-Antenna Perspective

    Zhiguo Ding, Robert Schober, H. Vincent Poor

    cs.ITeess.SParXiv:2412.02376v12024
  29. LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

    Joon Sung Park, Carolyn Q. Zou, Jonne Kamphorst +8

    cs.AIcs.HCcs.LGarXiv:2411.10109v32024
  30. Taming Visually Guided Sound Generation

    Vladimir Iashin, Esa Rahtu

    cs.CVcs.AIcs.LGarXiv:2110.08791v12021
  31. From PINNs to PIKANs: Recent Advances in Physics-Informed Machine Learning

    Juan Diego Toscano, Vivek Oommen, Alan John Varghese +4

    cs.LGcs.AIphysics.comp-pharXiv:2410.13228v22024
  32. Fast-dLLM v2: Efficient Block-Diffusion LLM

    Chengyue Wu, Hao Zhang, Shuchen Xue +7

    cs.CLarXiv:2509.26328v12025
  33. A Survey on Diffusion Models for Inverse Problems

    Giannis Daras, Hyungjin Chung, Chieh-Hsin Lai +5

    cs.LGcs.AIcs.CVarXiv:2410.00083v12024
  34. A Framework for the Robust Evaluation of Sound Event Detection

    Cagdas Bilen, Giacomo Ferroni, Francesco Tuveri +2

    eess.AScs.SDarXiv:1910.08440v22019
  35. Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

    Zhangchen Xu, Fengqing Jiang, Luyao Niu +4

    cs.CLcs.AIarXiv:2406.08464v22024
    Summaries:한국어
  36. Simple and Effective Masked Diffusion Language Models

    Subham Sekhar Sahoo, Marianne Arriola, Yair Schiff +5

    cs.CLcs.AIcs.LGarXiv:2406.07524v22024
  37. InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation

    Haofan Wang, Matteo Spinelli, Qixun Wang +3

    cs.CVarXiv:2404.02733v22024
  38. CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification

    Hanrong Zhang, Shicheng Fan, Henry Peng Zou +11

    cs.AIarXiv:2604.01687v32026
  39. Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

    Luca Soldaini, Rodney Kinney, Akshita Bhagia +33

    cs.CLarXiv:2402.00159v22024
  40. RewardBench: Evaluating Reward Models for Language Modeling

    Nathan Lambert, Valentina Pyatkin, Jacob Morrison +9

    cs.LGarXiv:2403.13787v22024
  41. Keeping it Simple: Language Models can learn Complex Molecular Distributions

    Daniel Flam-Shepherd, Kevin Zhu, Alán Aspuru-Guzik

    cs.LGcs.AIq-bio.QMarXiv:2112.03041v12021
  42. AnalysisBank: An Expert Analysis Pattern Library for Financial Report Generation

    Yajing Yang, Yunshan Ma, Kelvin J. L. Koa +1

    cs.AIarXiv:2609.00818v12026
  43. Artificial Intelligence for Literature Reviews: Opportunities and Challenges

    Francisco Bolanos, Angelo Salatino, Francesco Osborne +1

    cs.AIcs.HCcs.IRarXiv:2402.08565v22024
  44. Rate Maximization for Downlink Pinching-Antenna Systems

    Yanqing Xu, Zhiguo Ding, George K. Karagiannidis

    cs.ITeess.SParXiv:2502.12629v12025
  45. Exploring consumers response to text-based chatbots in e-commerce: The moderating role of task complexity and chatbot disclosure

    Xusen Cheng, Ying Bao, Alex Zarifis +2

    cs.AIarXiv:2401.12247v12024
  46. Rethinking FID: Towards a Better Evaluation Metric for Image Generation

    Sadeep Jayasumana, Srikumar Ramalingam, Andreas Veit +3

    cs.CVarXiv:2401.09603v22023
  47. Momentum Improves Normalized SGD

    Ashok Cutkosky, Harsh Mehta

    cs.LGmath.OCstat.MLarXiv:2002.03305v22020
  48. Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs

    Shengbang Tong, Zhuang Liu, Yuexiang Zhai +3

    cs.CVarXiv:2401.06209v22024
  49. Hiding Among the Clones: A Simple and Nearly Optimal Analysis of Privacy Amplification by Shuffling

    Vitaly Feldman, Audra McMillan, Kunal Talwar

    cs.LGcs.CRcs.DSarXiv:2012.12803v32020
  50. Task-Oriented Dialog Systems that Consider Multiple Appropriate Responses under the Same Context

    Yichi Zhang, Zhijian Ou, Zhou Yu

    cs.CLcs.AIarXiv:1911.10484v22019
  51. Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs

    Seungwoo Jung, Dohyeok Kwon, Seungmin Cha +4

    cs.AIarXiv:2609.00575v22026
  52. Empowering Edge Intelligence: A Comprehensive Survey on On-Device AI Models

    Xubin Wang, Zhiqing Tang, Jianxiong Guo +4

    cs.AIcs.LGcs.NIarXiv:2503.06027v22025
  53. Demonstration of fidelity improvement using dynamical decoupling with superconducting qubits

    Bibek Pokharel, Namit Anand, Benjamin Fortman +1

    quant-pharXiv:1807.08768v22018
  54. Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

    Kristen Grauman, Andrew Westbury, Lorenzo Torresani +98

    cs.CVcs.AIarXiv:2311.18259v42023
  55. Further Remarks on Separating Words

    John Nicol

    cs.FLarXiv:2608.30928v12026
  56. FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving

    Shuang Zeng, Xinyuan Chang, Mengwei Xie +6

    cs.CVarXiv:2505.17685v32025
  57. Point Transformer V3: Simpler, Faster, Stronger

    Xiaoyang Wu, Li Jiang, Peng-Shuai Wang +6

    cs.CVarXiv:2312.10035v22023
  58. Photorealistic Video Generation with Diffusion Models

    Agrim Gupta, Lijun Yu, Kihyuk Sohn +6

    cs.CVcs.AIcs.LGarXiv:2312.06662v12023
  59. The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning

    Bill Yuchen Lin, Abhilasha Ravichander, Ximing Lu +5

    cs.CLcs.AIarXiv:2312.01552v12023
  60. Generative Adversarial Networks: A Survey Towards Private and Secure Applications

    Zhipeng Cai, Zuobin Xiong, Honghui Xu +3

    cs.LGcs.CRarXiv:2106.03785v12021