Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
59,281 to 59,340 of 61,306
HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining
Juncheng Ma, Jianxin Bi, Yufan Deng +19
cs.CVarXiv:2606.20521v12026DF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis
Cheng-You Lu, Yi-Shan Hung, Wei-Ling Chi +6
cs.CVcs.AIarXiv:2604.13416v32026Beyond Static Leaderboards: Predictive Validity for the Evaluation of LLM Agents
Dhaval C. Patel, Kaoutar El Maghraoui, Shuxin Lin +58
cs.AIarXiv:2606.19704v12026When, Where, and How: Adaptive Binning for Tabular Self-Supervised Learning
Daehwan Kim, Haejun Chung, Ikbeom Jang
cs.LGcs.AIarXiv:2606.19827v12026Toward Parking Spot Occupancy Recognition: A Self-Supervised Approach
Luan Marko Kujavski, Rayson Laroca, Paulo Lisboa de Almeida
cs.CVarXiv:2606.20886v12026Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention
Vishesh Tripathi, Abhay Kumar
cs.LGarXiv:2606.20945v22026QG-MIL: A Gated Transformer Aggregator for Domain-Agnostic Multiple Instance Learning in Medical Imaging
Luca Zedda, Davide Antonio Mura, Cecilia Di Ruberto +4
cs.CVarXiv:2606.20027v12026HydraHead: From Head-Level Functional Heterogeneity to Specialized Attention Hybridization
Zhentao Tan, Wei Chen, Jingyi Shen +4
cs.CLarXiv:2606.20097v12026An Exploratory Case Study of LLM-Assisted Refactoring and Gameplay Feature Generation in an Endless Runner Game
Jan Wunderlich, Markus Kleffmann, Sebastian Lempert
cs.SEcs.AIarXiv:2606.21171v12026EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos
Yifan Zhong, Zhang Chen, Tianrui Guan +13
cs.ROarXiv:2607.09701v12026Causal Discovery in the Era of Agents
Yujia Zheng, Vishal Verma, Mantej Gill +3
cs.AIcs.LGcs.SEarXiv:2606.23608v12026Arbor: Explicit Geometric Conditioning for Controllable 3D Asset Generation
Jan-Niklas Dihlmann, Andreas Engelhardt, Simon Donne +2
cs.CVcs.GRarXiv:2606.23514v12026Vera: A Layered Diffusion Model for Content-Preserving Video Editing
Hongkai Zheng, Ta-Ying Cheng, Benjamin Klein +2
cs.CVarXiv:2606.23610v12026EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions
Jincheng Zhong, Weizhi Wang, Che Jiang +5
cs.CLcs.SEarXiv:2606.23654v12026PhoneBuddy: Training Open Models for Agentic Phone Use
Zhengyang Tang, Xin Lai, Pengyuan Lyu +23
cs.CLcs.AIarXiv:2606.23049v22026ChartWalker: Benchmarking the Cross-Chart RAG Task with Hierarchical Knowledge Graphs
Ning Tang, Chenghan Xie, Hanyang Yuan +6
cs.IRarXiv:2606.23997v12026VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct
Haoling Li, Kai Zheng, Jie Wu +4
cs.AIcs.CLcs.CVarXiv:2606.23543v12026Multi4D: High-Fidelity Dynamic Gaussian Splatting via Multi-Level Competitive Allocation
Rui Wang, Quentin Lohmeyer, Siyu Tang +1
cs.CVarXiv:2606.22197v12026OpenBioRQ: Unsolved Biomedical Research Questions for Agents
Minbyul Jeong
cs.CLarXiv:2606.21959v12026Lexical Consensus: Grounded Word Learning and Shared Meaning in Artificial Agents
Patricio M. Vera
cs.CLcs.AIarXiv:2606.22207v12026Deeper is Not Always Better: Mitigating the Alignment Tax via Confident Layer Decoding
Xuanming Zhang, Sining Zhoubian, Yuxuan Chen +8
cs.CLarXiv:2606.21906v12026Interleaved Speech Language Models Latently Work In Text
Talia Sternberg, Gallil Maimon, Yossi Adi
cs.CLcs.LGcs.SDarXiv:2606.22473v12026Sapiens2
Rawal Khirodkar, He Wen, Julieta Martinez +3
cs.CVarXiv:2604.21681v12026MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation
Dongxia Liu, Jie Ma, Xiaochen Yang +7
cs.GRcs.CVcs.LGarXiv:2605.13857v12026Realiz3D: 3D Generation Made Photorealistic via Domain-Aware Learning
Ido Sobol, Kihyuk Sohn, Yoav Blum +4
cs.GRcs.CVcs.LGarXiv:2605.13852v12026IAM: Identity-Aware Human Motion and Shape Joint Generation
Wenqi Jia, Zekun Li, Abhay Mittal +6
cs.CVarXiv:2604.25164v12026V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think
Bingda Tang, Yuhui Zhang, Xiaohan Wang +3
cs.LGcs.CVarXiv:2604.23380v12026The Signs Were Always There: Training-Free Concept Detection and Steering in Raw Transformer Dimensions
Varun Reddy Nalagatla
cs.LGcs.AIarXiv:2606.12629v32026Intelligent Base Station Deployment in Urban Wireless Networks: A Geographic Data-Informed Digital Twin Approach
Zhenyu Tao, Yuxuan Li, Wei Xu +2
cs.NIcs.AIarXiv:2608.14599v12026Counsel: A Meta-Evaluation Dataset for Agentic Tasks
Sashank Pisupati, Henry Broomfield, Eujeong Choi +5
cs.AIcs.LGarXiv:2606.21627v12026Toward Open Weight Models Without Risks: Separating Public and Private Capabilities in LLMs
Charbel El Feghali, Arkil Patel, Nicholas Meade +3
cs.CRcs.CLarXiv:2606.21638v12026PoLAR: Factorizing Extent and Mode in Latent Actions for Robot Policy Learning
Youngjoon Jeong, Jihwan Yu, Minsoo Jo +2
cs.ROcs.AIcs.LGarXiv:2606.21139v12026From Entity Mentions to Tone: An LLM-Based Pipeline for Media Bias Analysis
Klesti Hoxha, Olti Qirici
cs.CLarXiv:2608.17454v12026PrivacyAlign: Contextual Privacy Alignment for LLM Agents
Manveer Singh Tamber, Abhay Puri, Marc-Etienne Brunet +3
cs.CLcs.AIcs.IRarXiv:2606.21710v12026Speaker Identity in Non-Verbal Vocalizations: Conditional Distillation and Mixture of Experts Approach
Tzu-Chieh Wei, Yi-Cheng Lin, Huang-Cheng Chou +4
eess.AScs.AIcs.SDarXiv:2606.21215v12026BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language
Qizhi Pei, Zhimeng Zhou, Yi Duan +9
cs.CLcs.AIcs.LGarXiv:2606.22138v12026Libretto: Giving LLM Agents a Sense of Musical Structure
Yichen Xu
cs.SDcs.AIarXiv:2606.22708v12026An Agentic Framework Using Rules and LLMs for Embedding and Annotating Descriptive Document Layouts: A Plant Science Use Case
Nicolas Turenne, Youcef Sklab, Eric Chenin +1
cs.AIarXiv:2608.14587v12026Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing
Anastasiia Filippova, David Grangier, Marco Cuturi +1
cs.LGcs.AIarXiv:2604.22782v12026Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO
Zhiyuan Zeng, Jiameng Huang, Zhangyue Yin +8
cs.LGcs.AIcs.CLarXiv:2605.04077v12026Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment
Bobo Li, Rui Wu, Zibo Ji +5
cs.CLcs.AIcs.CYarXiv:2604.19548v12026TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction
Chengye Wang, Lin Fu, Zexi Kuang +1
cs.CLarXiv:2604.22880v12026Credal Concept Bottleneck Models for Epistemic-Aleatoric Uncertainty Decomposition
Tanmoy Mukherjee, Thomas Bailleux, Pierre Marquis +1
cs.AIarXiv:2604.24170v12026Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling
Zhen Ye, Xu Tan, Aoxiong Yin +8
cs.CVcs.CLcs.MMarXiv:2604.23586v22026Fool's Gold: Defensive Deception Against Safety-Removal Attacks on Open-Weight Models
Mark Russinovich
cs.AIcs.CRarXiv:2608.17202v12026S2-MoE: Enabling Efficient Self-Speculative Decoding for Mixture-of-Experts on Edge Devices
Haochen Huang, Shengxuan Qiu, Meng Li
cs.AIarXiv:2608.15018v12026aDSL: Agentic 3D Creation via Joint Agent-Program Design
Rui-Huan Wang, Si-Tong Wei, Jia-Qi He +3
cs.GRcs.CVarXiv:2608.17975v12026Toward Safe LLM Agents: A Survey of Specification, Verification, and Enforcement
Pierre Dantas, Lucas Cordeiro, Ehsan Nowroozi +1
cs.AIarXiv:2608.14590v120266G Native AI and Channel Foundation Models
Shugong Xu, Jun Jiang, Yuan Gao
eess.SPcs.ITcs.LGarXiv:2608.14591v12026Geometry Is Not Robustness: A Trajectory-Level Study of PGD Evaluation
Dhairysheel Durgule
cs.LGarXiv:2608.14594v12026iTryOn: Mastering Interactive Video Virtual Try-On with Spatial-Semantic Guidance
Jun Zheng, Zhengze Xu, Mengting Chen +6
cs.CVarXiv:2605.21431v22026StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding
Keming Wu, Baoyi Wang, Kaichen Zhang +7
cs.CVarXiv:2608.16320v12026Summaries:한국어LiveHouse-TS: An Open-world Living Benchmark for Time Series Foundation Models
Haomin Wen, Ziyu Zhou, Qingxiang Liu +2
cs.AIarXiv:2608.17299v12026Characterizing Narrative Content in Web-scale LLM Pretraining Data
Teagan Johnson, Elliott Ash, Andrew Piper +1
cs.CLarXiv:2606.19468v12026UniDot: A Unified Network for Sequence Modeling and Feature Interaction in Large-scale Recommendation
Rongcheng Lin, Yan Sun, Jamey Zhang +4
cs.IRcs.AIarXiv:2608.16797v12026Skill2Query: Exploiting Skill Structure to Generate Pseudo-Queries for Agent Skill Retrieval
Lihui Ding, Zihan Guo, Bingwei Lu +5
cs.CLcs.IRarXiv:2608.16071v12026A Theoretical Framework for Parallel Lifelong MAPF Using Group Decentralized Planning
Alex DeWeese, Jiaoyang Li, Guannan Qu
cs.MAcs.AIcs.ROarXiv:2608.17928v12026Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals
Joao Fonseca, Rodrigo Rodrigues, Paolo Romano
cs.AIcs.LGarXiv:2608.17687v12026Analysis of Types of Inquiries in Student-AI Interaction: A case study of two CS2 tasks
Matin Amoozadeh, Amin Alipour
cs.HCcs.AIarXiv:2608.17919v12026Learnware for CSI Feedback: Scene-specific Small Models Can Do Big
Xiangyi Li, Jiajia Guo, Chao-Kai Wen +3
cs.ITcs.AIeess.SParXiv:2608.17760v12026