Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

57,781 to 57,840 of 61,351

  1. Unsupervised Visual Representation Learning by Context Prediction

    Carl Doersch, Abhinav Gupta, Alexei A. Efros

    cs.CVarXiv:1505.05192v32015
  2. Evaluating Music Context Preservation: A Multi-facet Framework for Music Editing Systems

    Yash Vishe, Eric Xue, Xunyi Jiang +4

    cs.SDcs.AIarXiv:2512.14629v22025
  3. Cyclical Learning Rates for Training Neural Networks

    Leslie N. Smith

    cs.CVcs.LGcs.NEarXiv:1506.01186v62015
  4. $τ_0$-VLA: a Hierarchical Robot Foundation Model with World-Model-Guided Test-Time Computation

    Xiaowei Cai, Yunuo Cai, Bingao Chen +36

    cs.ROarXiv:2608.16885v12026
  5. Improving Neural Machine Translation Models with Monolingual Data

    Rico Sennrich, Barry Haddow, Alexandra Birch

    cs.CLarXiv:1511.06709v42015
  6. PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents

    Seongjae Kang, Taehyung Yu, Sung Ju Hwang

    cs.AIcs.CLcs.LGarXiv:2608.19861v12026
  7. Session-based Recommendations with Recurrent Neural Networks

    Balázs Hidasi, Alexandros Karatzoglou, Linas Baltrunas +1

    cs.LGcs.IRcs.NEarXiv:1511.06939v42015
  8. FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skills

    Zeyu Ren, Ling Yue, Ran Li +5

    cs.AIarXiv:2607.21596v22026
  9. PaLM-E: An Embodied Multimodal Language Model

    Danny Driess, Fei Xia, Mehdi S. M. Sajjadi +19

    cs.LGcs.AIcs.ROarXiv:2303.03378v12023
  10. EXIMO: VLM Guided Exploration of VLA Policies

    Bhavya Sukhija, Oliver Groth, Mohit Shridhar +5

    cs.AIarXiv:2608.19891v12026
  11. DOTA: A Large-scale Dataset for Object Detection in Aerial Images

    Gui-Song Xia, Xiang Bai, Jian Ding +6

    cs.CVarXiv:1711.10398v32017
  12. Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners

    Umberto Cappellazzo, Xubo Liu, Stavros Petridis +1

    eess.AScs.AIcs.SDarXiv:2608.19863v12026
  13. Cross-lingual Language Model Pretraining

    Guillaume Lample, Alexis Conneau

    cs.CLarXiv:1901.07291v12019
  14. Big Bird: Transformers for Longer Sequences

    Manzil Zaheer, Guru Guruganesh, Avinava Dubey +8

    cs.LGcs.CLstat.MLarXiv:2007.14062v22020
  15. NTU RGB+D: A Large Scale Dataset for 3D Human Activity Analysis

    Amir Shahroudy, Jun Liu, Tian-Tsong Ng +1

    cs.CVarXiv:1604.02808v12016
  16. FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM Serving

    Qihang Fan, Huaibo Huang, Zhiying Wu +2

    cs.CLarXiv:2608.19758v12026
  17. A Large Dataset to Train Convolutional Networks for Disparity, Optical Flow, and Scene Flow Estimation

    Nikolaus Mayer, Eddy Ilg, Philip Häusser +4

    cs.CVcs.LGstat.MLarXiv:1512.02134v12015
  18. A Survey on Multi-Task Learning

    Yu Zhang, Qiang Yang

    cs.LGcs.AIarXiv:1707.08114v32017
  19. Fine-Grained Visual Classification of Aircraft

    Subhransu Maji, Esa Rahtu, Juho Kannala +2

    cs.CVarXiv:1306.5151v12013
  20. SWE-bench Science: Can Coding Agents Resolve Engineering Tasks in Science?

    Zhipeng Xu, Jiahao Lu, Yining Zheng +2

    cs.CLcs.SEarXiv:2608.19799v12026
  21. Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

    Mohammad Shoeybi, Mostofa Patwary, Raul Puri +3

    cs.CLarXiv:1909.08053v42019
  22. Locating and Editing Factual Associations in GPT

    Kevin Meng, David Bau, Alex Andonian +1

    cs.CLcs.LGarXiv:2202.05262v52022
  23. WithEveryone: Unified Planning and Identity Grounding for Group Image Generation

    Hengyuan Xu, Qixun Wang, Yiji Cheng +5

    cs.CVarXiv:2608.20336v12026
  24. CCNet: Criss-Cross Attention for Semantic Segmentation

    Zilong Huang, Xinggang Wang, Yunchao Wei +4

    cs.CVarXiv:1811.11721v22018
  25. Estimation and Inference of Heterogeneous Treatment Effects using Random Forests

    Stefan Wager, Susan Athey

    stat.MEmath.STstat.MLarXiv:1510.04342v42015
    Summaries:한국어
  26. FLOPs vs Real Work: The Importance of Replication in AI Efficiency Assessment

    Enrique Barba Roque, Luís Cruz

    cs.AIcs.PFarXiv:2608.14550v12026
  27. Self-Evolving Visual Questioner

    Yijun Liang, Hengguang Zhou, Ming Li +3

    cs.CVcs.LGarXiv:2606.13929v12026
  28. MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

    Qiming Shi, Yulong Tao, Linbo Jin +10

    cs.AIarXiv:2607.28956v22026
  29. The Authority Resolution Framework: A Five-Domain Ontology for Governing Who and What Decides, at Scale

    Parviz Shariff

    cs.AIarXiv:2608.15832v12026
  30. A survey of cross-validation procedures for model selection

    Sylvain Arlot, Alain Celisse

    math.STstat.APstat.MEarXiv:0907.4728v12009
  31. Figurative and Cultural Knowledge in LLMs: Investigating Cross-Domain Transfer through Fine-Tuning

    Mena Attia, Mona Diab, Thamar Solorio

    cs.CLarXiv:2608.18361v12026
  32. Privacy-Preserving Dataset Curation for Kuala Lumpur Urban Traffic: Grounded Vision-Language Detection with Spatial Vehicle-Context Filtering

    Mohammed Abdul Al Arafat Tanzin, Rudzidatul Akmam Dziyauddin

    cs.CVcs.AIcs.LGarXiv:2608.14724v12026
  33. MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications

    Andrew G. Howard, Menglong Zhu, Bo Chen +5

    cs.CVarXiv:1704.04861v12017
  34. Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

    Michael Ahn, Anthony Brohan, Noah Brown +42

    cs.ROcs.CLcs.LGarXiv:2204.01691v22022
  35. Attention U-Net: Learning Where to Look for the Pancreas

    Ozan Oktay, Jo Schlemper, Loic Le Folgoc +9

    cs.CVarXiv:1804.03999v32018
  36. What Makes Software Issue Resolution Tasks Difficult for Agents?

    Ebtesam Al-Haque, Brittany Johnson

    cs.SEcs.AIcs.CLarXiv:2608.18280v12026
  37. ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration

    Ruofeng Yang, Yongcan Li, Shuai Li

    cs.SEcs.AIarXiv:2605.03042v12026
  38. A Pre-Specified Construction-Confirmation Test of Operation-Level Causal Transfer Across Finite Isomorphic Symbolic Domains

    Xinyi Shan

    cs.LGarXiv:2608.15809v12026
  39. Towards a Physics Foundation Model

    Florian Wiesner, Zoë J. Gray, Matthias Wessling +1

    cs.LGcs.AIstat.MLarXiv:2509.13805v42025
  40. LUNG-KGMM: Knowledge-Guided Multimodal Learning for Lung Cancer Incidence Prediction

    Chunlei Yang, Shuyan Li, Zhong Cao

    cs.LGcs.CVarXiv:2608.14657v12026
    Summaries:한국어
  41. Xception: Deep Learning with Depthwise Separable Convolutions

    François Chollet

    cs.CVarXiv:1610.02357v32016
  42. When Single-Dataset Conclusions Fail: A 45-Task Study of Threshold Tuning and Resampling for Imbalanced Classification

    Diyorbek Musaev

    cs.AIcs.LGarXiv:2608.16147v12026
    Summaries:한국어
  43. Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning

    Weixin Liang, Yuhui Zhang, Yongchan Kwon +2

    cs.CLcs.AIcs.CVarXiv:2203.02053v22022
    Summaries:한국어
  44. Proactive Road Safety Intervention in Australia: Predicting Risky Driving Hotspots from Connected Vehicle Data

    Adriana-Simona Mihăiţă, Clarence Cheung, Artur Grigorev +2

    cs.LGcs.CYarXiv:2608.16913v12026
  45. Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models

    Aswin RRV, Jacob Dineen, Divij Handa +4

    cs.AIarXiv:2605.08472v12026
  46. Can LLMs Introspect? A Reality Check

    Shashwat Singh, Tal Linzen, Shauli Ravfogel

    cs.AIarXiv:2605.26242v12026
  47. Stable-Layers: Fine-Tuning Image Layer Decomposition Models with VLM-Scored Reinforcement Learning

    Ciara Rowles, Reshinth Adithyan, Nikhil Pinnaparaju +2

    cs.CVarXiv:2605.30257v12026
  48. Deep learning with convolutional neural networks for EEG decoding and visualization

    Robin Tibor Schirrmeister, Jost Tobias Springenberg, Lukas Dominique Josef Fiederer +6

    cs.LGcs.NEarXiv:1703.05051v52017
  49. The Astropy Project: Sustaining and Growing a Community-oriented Open-source Project and the Latest Major Release (v5.0) of the Core Package

    The Astropy Collaboration, Adrian M. Price-Whelan, Pey Lian Lim +133

    astro-ph.IMarXiv:2206.14220v12022
  50. Skill Blocks: How Should an Agent Load Its Skill? A Caching-Correct Comparison of Pre-load, On-Demand Tool-Loading, Progressive Disclosure, and Hybrid

    Hironobu Nakasuji

    cs.AIarXiv:2608.14943v12026
  51. MemFuse: Multi-Source Memory Fusion from Fragmented Observations

    Chao Li, Yuanfa Li, Wenhao Wu +3

    cs.CLcs.AIarXiv:2608.18704v12026
  52. GRNEdit: Efficient General Video Editing from a New Binary-Evidence Perspective in Generative Refinement Networks

    Feng Xie, Jiagao Hu, Fuhao Li +5

    cs.CVarXiv:2608.16328v12026
  53. ASI-Bench: At the Dawn of Artificial Superintelligence

    Junwei Zhou, Zhen Sun, Binyu Li +39

    cs.AIarXiv:2608.17271v12026
    Summaries:한국어
  54. $R^3$-Bench: LLMs Struggle with Resource-Rational Reasoning under Shared Budgets

    Peisong Wang, Zhiwei Ma, Bowen Liu +6

    cs.CLarXiv:2608.16033v12026
    Summaries:한국어
  55. Token Distribution versus Data Volume: Domain Balancing in Multi-Domain Meeting Summarisation

    Ashima Sood, Bryan Gardiner, Joan Condell

    cs.CLarXiv:2608.15935v12026
  56. Algorithm-Architecture Co-Design for Efficient VLA Inference via Speculative Inference and Verification

    Chunyu Qi, Zhuoran Song, Jian Weng +6

    cs.ROcs.AIarXiv:2608.15636v12026
  57. VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding

    Fan Zhang, Guangming Yao, Jinyang Wu +6

    cs.CVcs.CLarXiv:2608.14718v12026
  58. PACE-Bench: Benchmarking Physics Adaptation via Code Evolution in Dynamic Environments

    Yuhao Zhan, Bingxiang He, Zecong Tang +1

    cs.AIarXiv:2608.14441v12026
  59. MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided Refinement

    Lushi Pu, Weiming Zhang, Xinheng Xie +7

    cs.AIcs.CLarXiv:2608.14221v12026
  60. MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation

    Youjun Zhao, Alex Warren, Gary K. L. Tam +1

    cs.CVcs.LGarXiv:2608.07463v12026