Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
57,781 to 57,840 of 61,351
Unsupervised Visual Representation Learning by Context Prediction
Carl Doersch, Abhinav Gupta, Alexei A. Efros
cs.CVarXiv:1505.05192v32015Evaluating Music Context Preservation: A Multi-facet Framework for Music Editing Systems
Yash Vishe, Eric Xue, Xunyi Jiang +4
cs.SDcs.AIarXiv:2512.14629v22025Cyclical Learning Rates for Training Neural Networks
Leslie N. Smith
cs.CVcs.LGcs.NEarXiv:1506.01186v62015$τ_0$-VLA: a Hierarchical Robot Foundation Model with World-Model-Guided Test-Time Computation
Xiaowei Cai, Yunuo Cai, Bingao Chen +36
cs.ROarXiv:2608.16885v12026Improving Neural Machine Translation Models with Monolingual Data
Rico Sennrich, Barry Haddow, Alexandra Birch
cs.CLarXiv:1511.06709v42015PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents
Seongjae Kang, Taehyung Yu, Sung Ju Hwang
cs.AIcs.CLcs.LGarXiv:2608.19861v12026Session-based Recommendations with Recurrent Neural Networks
Balázs Hidasi, Alexandros Karatzoglou, Linas Baltrunas +1
cs.LGcs.IRcs.NEarXiv:1511.06939v42015FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skills
Zeyu Ren, Ling Yue, Ran Li +5
cs.AIarXiv:2607.21596v22026PaLM-E: An Embodied Multimodal Language Model
Danny Driess, Fei Xia, Mehdi S. M. Sajjadi +19
cs.LGcs.AIcs.ROarXiv:2303.03378v12023EXIMO: VLM Guided Exploration of VLA Policies
Bhavya Sukhija, Oliver Groth, Mohit Shridhar +5
cs.AIarXiv:2608.19891v12026DOTA: A Large-scale Dataset for Object Detection in Aerial Images
Gui-Song Xia, Xiang Bai, Jian Ding +6
cs.CVarXiv:1711.10398v32017Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners
Umberto Cappellazzo, Xubo Liu, Stavros Petridis +1
eess.AScs.AIcs.SDarXiv:2608.19863v12026Cross-lingual Language Model Pretraining
Guillaume Lample, Alexis Conneau
cs.CLarXiv:1901.07291v12019Big Bird: Transformers for Longer Sequences
Manzil Zaheer, Guru Guruganesh, Avinava Dubey +8
cs.LGcs.CLstat.MLarXiv:2007.14062v22020NTU RGB+D: A Large Scale Dataset for 3D Human Activity Analysis
Amir Shahroudy, Jun Liu, Tian-Tsong Ng +1
cs.CVarXiv:1604.02808v12016FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM Serving
Qihang Fan, Huaibo Huang, Zhiying Wu +2
cs.CLarXiv:2608.19758v12026A Large Dataset to Train Convolutional Networks for Disparity, Optical Flow, and Scene Flow Estimation
Nikolaus Mayer, Eddy Ilg, Philip Häusser +4
cs.CVcs.LGstat.MLarXiv:1512.02134v12015A Survey on Multi-Task Learning
Yu Zhang, Qiang Yang
cs.LGcs.AIarXiv:1707.08114v32017Fine-Grained Visual Classification of Aircraft
Subhransu Maji, Esa Rahtu, Juho Kannala +2
cs.CVarXiv:1306.5151v12013SWE-bench Science: Can Coding Agents Resolve Engineering Tasks in Science?
Zhipeng Xu, Jiahao Lu, Yining Zheng +2
cs.CLcs.SEarXiv:2608.19799v12026Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
Mohammad Shoeybi, Mostofa Patwary, Raul Puri +3
cs.CLarXiv:1909.08053v42019Locating and Editing Factual Associations in GPT
Kevin Meng, David Bau, Alex Andonian +1
cs.CLcs.LGarXiv:2202.05262v52022WithEveryone: Unified Planning and Identity Grounding for Group Image Generation
Hengyuan Xu, Qixun Wang, Yiji Cheng +5
cs.CVarXiv:2608.20336v12026CCNet: Criss-Cross Attention for Semantic Segmentation
Zilong Huang, Xinggang Wang, Yunchao Wei +4
cs.CVarXiv:1811.11721v22018Estimation and Inference of Heterogeneous Treatment Effects using Random Forests
Stefan Wager, Susan Athey
stat.MEmath.STstat.MLarXiv:1510.04342v42015Summaries:한국어FLOPs vs Real Work: The Importance of Replication in AI Efficiency Assessment
Enrique Barba Roque, Luís Cruz
cs.AIcs.PFarXiv:2608.14550v12026Self-Evolving Visual Questioner
Yijun Liang, Hengguang Zhou, Ming Li +3
cs.CVcs.LGarXiv:2606.13929v12026MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations
Qiming Shi, Yulong Tao, Linbo Jin +10
cs.AIarXiv:2607.28956v22026The Authority Resolution Framework: A Five-Domain Ontology for Governing Who and What Decides, at Scale
Parviz Shariff
cs.AIarXiv:2608.15832v12026A survey of cross-validation procedures for model selection
Sylvain Arlot, Alain Celisse
math.STstat.APstat.MEarXiv:0907.4728v12009Figurative and Cultural Knowledge in LLMs: Investigating Cross-Domain Transfer through Fine-Tuning
Mena Attia, Mona Diab, Thamar Solorio
cs.CLarXiv:2608.18361v12026Privacy-Preserving Dataset Curation for Kuala Lumpur Urban Traffic: Grounded Vision-Language Detection with Spatial Vehicle-Context Filtering
Mohammed Abdul Al Arafat Tanzin, Rudzidatul Akmam Dziyauddin
cs.CVcs.AIcs.LGarXiv:2608.14724v12026MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications
Andrew G. Howard, Menglong Zhu, Bo Chen +5
cs.CVarXiv:1704.04861v12017Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Michael Ahn, Anthony Brohan, Noah Brown +42
cs.ROcs.CLcs.LGarXiv:2204.01691v22022Attention U-Net: Learning Where to Look for the Pancreas
Ozan Oktay, Jo Schlemper, Loic Le Folgoc +9
cs.CVarXiv:1804.03999v32018What Makes Software Issue Resolution Tasks Difficult for Agents?
Ebtesam Al-Haque, Brittany Johnson
cs.SEcs.AIcs.CLarXiv:2608.18280v12026ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration
Ruofeng Yang, Yongcan Li, Shuai Li
cs.SEcs.AIarXiv:2605.03042v12026A Pre-Specified Construction-Confirmation Test of Operation-Level Causal Transfer Across Finite Isomorphic Symbolic Domains
Xinyi Shan
cs.LGarXiv:2608.15809v12026Towards a Physics Foundation Model
Florian Wiesner, Zoë J. Gray, Matthias Wessling +1
cs.LGcs.AIstat.MLarXiv:2509.13805v42025LUNG-KGMM: Knowledge-Guided Multimodal Learning for Lung Cancer Incidence Prediction
Chunlei Yang, Shuyan Li, Zhong Cao
cs.LGcs.CVarXiv:2608.14657v12026Summaries:한국어Xception: Deep Learning with Depthwise Separable Convolutions
François Chollet
cs.CVarXiv:1610.02357v32016When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification
Diyorbek Musaev
cs.AIcs.LGarXiv:2608.16147v12026Summaries:한국어Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning
Weixin Liang, Yuhui Zhang, Yongchan Kwon +2
cs.CLcs.AIcs.CVarXiv:2203.02053v22022Summaries:한국어Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data
Adriana-Simona Mihăiţă, Clarence Cheung, Artur Grigorev +2
cs.LGcs.CYarXiv:2608.16913v12026Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models
Aswin RRV, Jacob Dineen, Divij Handa +4
cs.AIarXiv:2605.08472v12026Can LLMs Introspect? A Reality Check
Shashwat Singh, Tal Linzen, Shauli Ravfogel
cs.AIarXiv:2605.26242v12026Stable-Layers: Fine-Tuning Image Layer Decomposition Models with VLM-Scored Reinforcement Learning
Ciara Rowles, Reshinth Adithyan, Nikhil Pinnaparaju +2
cs.CVarXiv:2605.30257v12026Deep learning with convolutional neural networks for EEG decoding and visualization
Robin Tibor Schirrmeister, Jost Tobias Springenberg, Lukas Dominique Josef Fiederer +6
cs.LGcs.NEarXiv:1703.05051v52017The Astropy Project: Sustaining and Growing a Community-oriented Open-source Project and the Latest Major Release (v5.0) of the Core Package
The Astropy Collaboration, Adrian M. Price-Whelan, Pey Lian Lim +133
astro-ph.IMarXiv:2206.14220v12022Skill Blocks: How Should an Agent Load Its Skill? A Caching-Correct Comparison of Pre-load, On-Demand Tool-Loading, Progressive Disclosure, and Hybrid
Hironobu Nakasuji
cs.AIarXiv:2608.14943v12026MemFuse: Multi-Source Memory Fusion from Fragmented Observations
Chao Li, Yuanfa Li, Wenhao Wu +3
cs.CLcs.AIarXiv:2608.18704v12026GRNEdit: Efficient General Video Editing from a New Binary-Evidence Perspective in Generative Refinement Networks
Feng Xie, Jiagao Hu, Fuhao Li +5
cs.CVarXiv:2608.16328v12026ASI-Bench: At the Dawn of Artificial Superintelligence
Junwei Zhou, Zhen Sun, Binyu Li +39
cs.AIarXiv:2608.17271v12026Summaries:한국어$R^3$-Bench: LLMs Struggle with Resource-Rational Reasoning under Shared Budgets
Peisong Wang, Zhiwei Ma, Bowen Liu +6
cs.CLarXiv:2608.16033v12026Summaries:한국어Token Distribution versus Data Volume: Domain Balancing in Multi-Domain Meeting Summarisation
Ashima Sood, Bryan Gardiner, Joan Condell
cs.CLarXiv:2608.15935v12026Algorithm-Architecture Co-Design for Efficient VLA Inference via Speculative Inference and Verification
Chunyu Qi, Zhuoran Song, Jian Weng +6
cs.ROcs.AIarXiv:2608.15636v12026VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding
Fan Zhang, Guangming Yao, Jinyang Wu +6
cs.CVcs.CLarXiv:2608.14718v12026PACE-Bench: Benchmarking Physics Adaptation via Code Evolution in Dynamic Environments
Yuhao Zhan, Bingxiang He, Zecong Tang +1
cs.AIarXiv:2608.14441v12026MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided Refinement
Lushi Pu, Weiming Zhang, Xinheng Xie +7
cs.AIcs.CLarXiv:2608.14221v12026MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation
Youjun Zhao, Alex Warren, Gary K. L. Tam +1
cs.CVcs.LGarXiv:2608.07463v12026