Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
22,261 to 22,320 of 61,134
Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions
Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
cs.CVcs.CLarXiv:1811.10652v32018No Language Left Behind: Scaling Human-Centered Machine Translation
NLLB Team, Marta R. Costa-jussà, James Cross +36
cs.CLcs.AIarXiv:2207.04672v32022LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
Xueyang Zhou, Yangming Xu, Guiyao Tie +5
cs.CVcs.ROarXiv:2510.03827v22025Spatial-Angular Interaction for Light Field Image Super-Resolution
Yingqian Wang, Longguang Wang, Jungang Yang +3
eess.IVcs.CVarXiv:1912.07849v32019Simple Contrastive Graph Clustering
Yue Liu, Xihong Yang, Sihang Zhou +1
cs.LGcs.AIarXiv:2205.07865v32022Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective
Kun Chen, Peng Shi, Fanfan Liu +4
cs.LGcs.AIcs.CLarXiv:2602.09782v22026Trust The Typical
Debargha Ganguly, Sreehari Sankar, Biyao Zhang +8
cs.CLcs.AIcs.DCarXiv:2602.04581v12026Does "AI" stand for augmenting inequality in the era of covid-19 healthcare?
David Leslie, Anjali Mazumder, Aidan Peppin +2
cs.CYcs.LGarXiv:2105.07844v12021Large Language Model Agent: A Survey on Methodology, Applications and Challenges
Junyu Luo, Weizhi Zhang, Ye Yuan +23
cs.CLarXiv:2503.21460v12025Agentic AI: A Comprehensive Survey of Architectures, Applications, and Future Directions
Mohamad Abou Ali, Fadi Dornaika
cs.AIcs.LGarXiv:2510.25445v12025Context Learning for Multi-Agent Discussion
Xingyuan Hua, Sheng Yue, Xinyi Li +3
cs.AIcs.LGcs.MAarXiv:2602.02350v32026Steering LLMs via Scalable Interactive Oversight
Enyu Zhou, Zhiheng Xi, Long Ma +9
cs.AIcs.LGarXiv:2602.04210v22026Sensor Fault Detection, Isolation and Identification Using Multiple Model-based Hybrid Kalman Filter for Gas Turbine Engines
Bahareh Pourbabaee, Nader Meskin, Khashayar Khorasani
eess.SYarXiv:1505.02063v22015Alternating Reinforcement Learning for Rubric-Based Reward Modeling in Non-Verifiable LLM Post-Training
Ran Xu, Tianci Liu, Zihan Dong +6
cs.CLcs.LGarXiv:2602.01511v22026Multi-agent Architecture Search via Agentic Supernet
Guibin Zhang, Luyang Niu, Junfeng Fang +3
cs.LGcs.CLcs.MAarXiv:2502.04180v22025FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
Ajay Patel, Colin Raffel, Chris Callison-Burch
cs.CLcs.LGarXiv:2601.22146v32026Ethics-Based Auditing to Develop Trustworthy AI
Jakob Mokander, Luciano Floridi
cs.CYcs.AIarXiv:2105.00002v12021Deep Learning Object Detection Methods for Ecological Camera Trap Data
Stefan Schneider, Graham W. Taylor, Stefan C. Kremer
cs.CVarXiv:1803.10842v12018VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agents
Zirui Wang, Junyi Zhang, Jiaxin Ge +9
cs.CVarXiv:2601.16973v12026RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete
Yuheng Ji, Huajie Tan, Jiayu Shi +14
cs.ROcs.CVarXiv:2502.21257v22025Your Brain on ChatGPT: Accumulation of Cognitive Debt when Using an AI Assistant for Essay Writing Task
Nataliya Kosmyna, Eugene Hauptmann, Ye Tong Yuan +5
cs.AIarXiv:2506.08872v22025InT: Self-Proposed Interventions Enable Credit Assignment in LLM Reasoning
Matthew Y. R. Yang, Hao Bai, Ian Wu +3
cs.LGcs.AIcs.CLarXiv:2601.14209v12026Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents
Vidya Srinivas, Zachary Englhardt, Vikram Iyer +1
cs.CLarXiv:2511.07397v32025Summaries:한국어A Generative Appearance Model for End-to-end Video Object Segmentation
Joakim Johnander, Martin Danelljan, Emil Brissman +2
cs.CVarXiv:1811.11611v22018A Survey on LLM-as-a-Judge
Jiawei Gu, Xuhui Jiang, Zhichao Shi +13
cs.CLcs.AIarXiv:2411.15594v62024An All-in-One Network for Dehazing and Beyond
Boyi Li, Xiulian Peng, Zhangyang Wang +2
cs.CVcs.AIarXiv:1707.06543v12017Introduction to Machine Learning
Laurent Younes
stat.MLcs.LGarXiv:2409.02668v22024Flexible-Antenna Systems: A Pinching-Antenna Perspective
Zhiguo Ding, Robert Schober, H. Vincent Poor
cs.ITeess.SParXiv:2412.02376v12024LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals
Joon Sung Park, Carolyn Q. Zou, Jonne Kamphorst +8
cs.AIcs.HCcs.LGarXiv:2411.10109v32024Taming Visually Guided Sound Generation
Vladimir Iashin, Esa Rahtu
cs.CVcs.AIcs.LGarXiv:2110.08791v12021From PINNs to PIKANs: Recent Advances in Physics-Informed Machine Learning
Juan Diego Toscano, Vivek Oommen, Alan John Varghese +4
cs.LGcs.AIphysics.comp-pharXiv:2410.13228v22024Fast-dLLM v2: Efficient Block-Diffusion LLM
Chengyue Wu, Hao Zhang, Shuchen Xue +7
cs.CLarXiv:2509.26328v12025A Survey on Diffusion Models for Inverse Problems
Giannis Daras, Hyungjin Chung, Chieh-Hsin Lai +5
cs.LGcs.AIcs.CVarXiv:2410.00083v12024A Framework for the Robust Evaluation of Sound Event Detection
Cagdas Bilen, Giacomo Ferroni, Francesco Tuveri +2
eess.AScs.SDarXiv:1910.08440v22019Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing
Zhangchen Xu, Fengqing Jiang, Luyao Niu +4
cs.CLcs.AIarXiv:2406.08464v22024Summaries:한국어Simple and Effective Masked Diffusion Language Models
Subham Sekhar Sahoo, Marianne Arriola, Yair Schiff +5
cs.CLcs.AIcs.LGarXiv:2406.07524v22024InstantStyle: Free Lunch towards Style-Preserving in Text-to-Image Generation
Haofan Wang, Matteo Spinelli, Qixun Wang +3
cs.CVarXiv:2404.02733v22024CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification
Hanrong Zhang, Shicheng Fan, Henry Peng Zou +11
cs.AIarXiv:2604.01687v32026Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
Luca Soldaini, Rodney Kinney, Akshita Bhagia +33
cs.CLarXiv:2402.00159v22024RewardBench: Evaluating Reward Models for Language Modeling
Nathan Lambert, Valentina Pyatkin, Jacob Morrison +9
cs.LGarXiv:2403.13787v22024Keeping it Simple: Language Models can learn Complex Molecular Distributions
Daniel Flam-Shepherd, Kevin Zhu, Alán Aspuru-Guzik
cs.LGcs.AIq-bio.QMarXiv:2112.03041v12021AnalysisBank: An Expert Analysis Pattern Library for Financial Report Generation
Yajing Yang, Yunshan Ma, Kelvin J. L. Koa +1
cs.AIarXiv:2609.00818v12026Artificial Intelligence for Literature Reviews: Opportunities and Challenges
Francisco Bolanos, Angelo Salatino, Francesco Osborne +1
cs.AIcs.HCcs.IRarXiv:2402.08565v22024Rate Maximization for Downlink Pinching-Antenna Systems
Yanqing Xu, Zhiguo Ding, George K. Karagiannidis
cs.ITeess.SParXiv:2502.12629v12025Exploring consumers response to text-based chatbots in e-commerce: The moderating role of task complexity and chatbot disclosure
Xusen Cheng, Ying Bao, Alex Zarifis +2
cs.AIarXiv:2401.12247v12024Rethinking FID: Towards a Better Evaluation Metric for Image Generation
Sadeep Jayasumana, Srikumar Ramalingam, Andreas Veit +3
cs.CVarXiv:2401.09603v22023Momentum Improves Normalized SGD
Ashok Cutkosky, Harsh Mehta
cs.LGmath.OCstat.MLarXiv:2002.03305v22020Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Shengbang Tong, Zhuang Liu, Yuexiang Zhai +3
cs.CVarXiv:2401.06209v22024Hiding Among the Clones: A Simple and Nearly Optimal Analysis of Privacy Amplification by Shuffling
Vitaly Feldman, Audra McMillan, Kunal Talwar
cs.LGcs.CRcs.DSarXiv:2012.12803v32020Task-Oriented Dialog Systems that Consider Multiple Appropriate Responses under the Same Context
Yichi Zhang, Zhijian Ou, Zhou Yu
cs.CLcs.AIarXiv:1911.10484v22019Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs
Seungwoo Jung, Dohyeok Kwon, Seungmin Cha +4
cs.AIarXiv:2609.00575v22026Empowering Edge Intelligence: A Comprehensive Survey on On-Device AI Models
Xubin Wang, Zhiqing Tang, Jianxiong Guo +4
cs.AIcs.LGcs.NIarXiv:2503.06027v22025Demonstration of fidelity improvement using dynamical decoupling with superconducting qubits
Bibek Pokharel, Namit Anand, Benjamin Fortman +1
quant-pharXiv:1807.08768v22018Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives
Kristen Grauman, Andrew Westbury, Lorenzo Torresani +98
cs.CVcs.AIarXiv:2311.18259v42023Further Remarks on Separating Words
John Nicol
cs.FLarXiv:2608.30928v12026FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
Shuang Zeng, Xinyuan Chang, Mengwei Xie +6
cs.CVarXiv:2505.17685v32025Point Transformer V3: Simpler, Faster, Stronger
Xiaoyang Wu, Li Jiang, Peng-Shuai Wang +6
cs.CVarXiv:2312.10035v22023Photorealistic Video Generation with Diffusion Models
Agrim Gupta, Lijun Yu, Kihyuk Sohn +6
cs.CVcs.AIcs.LGarXiv:2312.06662v12023The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning
Bill Yuchen Lin, Abhilasha Ravichander, Ximing Lu +5
cs.CLcs.AIarXiv:2312.01552v12023Generative Adversarial Networks: A Survey Towards Private and Secure Applications
Zhipeng Cai, Zuobin Xiong, Honghui Xu +3
cs.LGcs.CRarXiv:2106.03785v12021