Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

59,281 to 59,340 of 61,306

  1. HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining

    Juncheng Ma, Jianxin Bi, Yufan Deng +19

    cs.CVarXiv:2606.20521v12026
  2. DF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis

    Cheng-You Lu, Yi-Shan Hung, Wei-Ling Chi +6

    cs.CVcs.AIarXiv:2604.13416v32026
  3. Beyond Static Leaderboards: Predictive Validity for the Evaluation of LLM Agents

    Dhaval C. Patel, Kaoutar El Maghraoui, Shuxin Lin +58

    cs.AIarXiv:2606.19704v12026
  4. When, Where, and How: Adaptive Binning for Tabular Self-Supervised Learning

    Daehwan Kim, Haejun Chung, Ikbeom Jang

    cs.LGcs.AIarXiv:2606.19827v12026
  5. Toward Parking Spot Occupancy Recognition: A Self-Supervised Approach

    Luan Marko Kujavski, Rayson Laroca, Paulo Lisboa de Almeida

    cs.CVarXiv:2606.20886v12026
  6. Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention

    Vishesh Tripathi, Abhay Kumar

    cs.LGarXiv:2606.20945v22026
  7. QG-MIL: A Gated Transformer Aggregator for Domain-Agnostic Multiple Instance Learning in Medical Imaging

    Luca Zedda, Davide Antonio Mura, Cecilia Di Ruberto +4

    cs.CVarXiv:2606.20027v12026
  8. HydraHead: From Head-Level Functional Heterogeneity to Specialized Attention Hybridization

    Zhentao Tan, Wei Chen, Jingyi Shen +4

    cs.CLarXiv:2606.20097v12026
  9. An Exploratory Case Study of LLM-Assisted Refactoring and Gameplay Feature Generation in an Endless Runner Game

    Jan Wunderlich, Markus Kleffmann, Sebastian Lempert

    cs.SEcs.AIarXiv:2606.21171v12026
  10. EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos

    Yifan Zhong, Zhang Chen, Tianrui Guan +13

    cs.ROarXiv:2607.09701v12026
  11. Causal Discovery in the Era of Agents

    Yujia Zheng, Vishal Verma, Mantej Gill +3

    cs.AIcs.LGcs.SEarXiv:2606.23608v12026
  12. Arbor: Explicit Geometric Conditioning for Controllable 3D Asset Generation

    Jan-Niklas Dihlmann, Andreas Engelhardt, Simon Donne +2

    cs.CVcs.GRarXiv:2606.23514v12026
  13. Vera: A Layered Diffusion Model for Content-Preserving Video Editing

    Hongkai Zheng, Ta-Ying Cheng, Benjamin Klein +2

    cs.CVarXiv:2606.23610v12026
  14. EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions

    Jincheng Zhong, Weizhi Wang, Che Jiang +5

    cs.CLcs.SEarXiv:2606.23654v12026
  15. PhoneBuddy: Training Open Models for Agentic Phone Use

    Zhengyang Tang, Xin Lai, Pengyuan Lyu +23

    cs.CLcs.AIarXiv:2606.23049v22026
  16. ChartWalker: Benchmarking the Cross-Chart RAG Task with Hierarchical Knowledge Graphs

    Ning Tang, Chenghan Xie, Hanyang Yuan +6

    cs.IRarXiv:2606.23997v12026
  17. VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct

    Haoling Li, Kai Zheng, Jie Wu +4

    cs.AIcs.CLcs.CVarXiv:2606.23543v12026
  18. Multi4D: High-Fidelity Dynamic Gaussian Splatting via Multi-Level Competitive Allocation

    Rui Wang, Quentin Lohmeyer, Siyu Tang +1

    cs.CVarXiv:2606.22197v12026
  19. OpenBioRQ: Unsolved Biomedical Research Questions for Agents

    Minbyul Jeong

    cs.CLarXiv:2606.21959v12026
  20. Lexical Consensus: Grounded Word Learning and Shared Meaning in Artificial Agents

    Patricio M. Vera

    cs.CLcs.AIarXiv:2606.22207v12026
  21. Deeper is Not Always Better: Mitigating the Alignment Tax via Confident Layer Decoding

    Xuanming Zhang, Sining Zhoubian, Yuxuan Chen +8

    cs.CLarXiv:2606.21906v12026
  22. Interleaved Speech Language Models Latently Work In Text

    Talia Sternberg, Gallil Maimon, Yossi Adi

    cs.CLcs.LGcs.SDarXiv:2606.22473v12026
  23. Sapiens2

    Rawal Khirodkar, He Wen, Julieta Martinez +3

    cs.CVarXiv:2604.21681v12026
  24. MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation

    Dongxia Liu, Jie Ma, Xiaochen Yang +7

    cs.GRcs.CVcs.LGarXiv:2605.13857v12026
  25. Realiz3D: 3D Generation Made Photorealistic via Domain-Aware Learning

    Ido Sobol, Kihyuk Sohn, Yoav Blum +4

    cs.GRcs.CVcs.LGarXiv:2605.13852v12026
  26. IAM: Identity-Aware Human Motion and Shape Joint Generation

    Wenqi Jia, Zekun Li, Abhay Mittal +6

    cs.CVarXiv:2604.25164v12026
  27. V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think

    Bingda Tang, Yuhui Zhang, Xiaohan Wang +3

    cs.LGcs.CVarXiv:2604.23380v12026
  28. The Signs Were Always There: Training-Free Concept Detection and Steering in Raw Transformer Dimensions

    Varun Reddy Nalagatla

    cs.LGcs.AIarXiv:2606.12629v32026
  29. Intelligent Base Station Deployment in Urban Wireless Networks: A Geographic Data-Informed Digital Twin Approach

    Zhenyu Tao, Yuxuan Li, Wei Xu +2

    cs.NIcs.AIarXiv:2608.14599v12026
  30. Counsel: A Meta-Evaluation Dataset for Agentic Tasks

    Sashank Pisupati, Henry Broomfield, Eujeong Choi +5

    cs.AIcs.LGarXiv:2606.21627v12026
  31. Toward Open Weight Models Without Risks: Separating Public and Private Capabilities in LLMs

    Charbel El Feghali, Arkil Patel, Nicholas Meade +3

    cs.CRcs.CLarXiv:2606.21638v12026
  32. PoLAR: Factorizing Extent and Mode in Latent Actions for Robot Policy Learning

    Youngjoon Jeong, Jihwan Yu, Minsoo Jo +2

    cs.ROcs.AIcs.LGarXiv:2606.21139v12026
  33. From Entity Mentions to Tone: An LLM-Based Pipeline for Media Bias Analysis

    Klesti Hoxha, Olti Qirici

    cs.CLarXiv:2608.17454v12026
  34. PrivacyAlign: Contextual Privacy Alignment for LLM Agents

    Manveer Singh Tamber, Abhay Puri, Marc-Etienne Brunet +3

    cs.CLcs.AIcs.IRarXiv:2606.21710v12026
  35. Speaker Identity in Non-Verbal Vocalizations: Conditional Distillation and Mixture of Experts Approach

    Tzu-Chieh Wei, Yi-Cheng Lin, Huang-Cheng Chou +4

    eess.AScs.AIcs.SDarXiv:2606.21215v12026
  36. BioMatrix: Towards a Comprehensive Biological Foundation Model Spanning the Modality Matrix of Sequences, Structures, and Language

    Qizhi Pei, Zhimeng Zhou, Yi Duan +9

    cs.CLcs.AIcs.LGarXiv:2606.22138v12026
  37. Libretto: Giving LLM Agents a Sense of Musical Structure

    Yichen Xu

    cs.SDcs.AIarXiv:2606.22708v12026
  38. An Agentic Framework Using Rules and LLMs for Embedding and Annotating Descriptive Document Layouts: A Plant Science Use Case

    Nicolas Turenne, Youcef Sklab, Eric Chenin +1

    cs.AIarXiv:2608.14587v12026
  39. Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing

    Anastasiia Filippova, David Grangier, Marco Cuturi +1

    cs.LGcs.AIarXiv:2604.22782v12026
  40. Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO

    Zhiyuan Zeng, Jiameng Huang, Zhangyue Yin +8

    cs.LGcs.AIcs.CLarXiv:2605.04077v12026
  41. Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

    Bobo Li, Rui Wu, Zibo Ji +5

    cs.CLcs.AIcs.CYarXiv:2604.19548v12026
  42. TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

    Chengye Wang, Lin Fu, Zexi Kuang +1

    cs.CLarXiv:2604.22880v12026
  43. Credal Concept Bottleneck Models for Epistemic-Aleatoric Uncertainty Decomposition

    Tanmoy Mukherjee, Thomas Bailleux, Pierre Marquis +1

    cs.AIarXiv:2604.24170v12026
  44. Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling

    Zhen Ye, Xu Tan, Aoxiong Yin +8

    cs.CVcs.CLcs.MMarXiv:2604.23586v22026
  45. Fool's Gold: Defensive Deception Against Safety-Removal Attacks on Open-Weight Models

    Mark Russinovich

    cs.AIcs.CRarXiv:2608.17202v12026
  46. S2-MoE: Enabling Efficient Self-Speculative Decoding for Mixture-of-Experts on Edge Devices

    Haochen Huang, Shengxuan Qiu, Meng Li

    cs.AIarXiv:2608.15018v12026
  47. aDSL: Agentic 3D Creation via Joint Agent-Program Design

    Rui-Huan Wang, Si-Tong Wei, Jia-Qi He +3

    cs.GRcs.CVarXiv:2608.17975v12026
  48. Toward Safe LLM Agents: A Survey of Specification, Verification, and Enforcement

    Pierre Dantas, Lucas Cordeiro, Ehsan Nowroozi +1

    cs.AIarXiv:2608.14590v12026
  49. 6G Native AI and Channel Foundation Models

    Shugong Xu, Jun Jiang, Yuan Gao

    eess.SPcs.ITcs.LGarXiv:2608.14591v12026
  50. Geometry Is Not Robustness: A Trajectory-Level Study of PGD Evaluation

    Dhairysheel Durgule

    cs.LGarXiv:2608.14594v12026
  51. iTryOn: Mastering Interactive Video Virtual Try-On with Spatial-Semantic Guidance

    Jun Zheng, Zhengze Xu, Mengting Chen +6

    cs.CVarXiv:2605.21431v22026
  52. StreamOPD: A Post-Training Recipe with Spatio-Temporal Cue Gating for Streaming Video Understanding

    Keming Wu, Baoyi Wang, Kaichen Zhang +7

    cs.CVarXiv:2608.16320v12026
    Summaries:한국어
  53. LiveHouse-TS: An Open-world Living Benchmark for Time Series Foundation Models

    Haomin Wen, Ziyu Zhou, Qingxiang Liu +2

    cs.AIarXiv:2608.17299v12026
  54. Characterizing Narrative Content in Web-scale LLM Pretraining Data

    Teagan Johnson, Elliott Ash, Andrew Piper +1

    cs.CLarXiv:2606.19468v12026
  55. UniDot: A Unified Network for Sequence Modeling and Feature Interaction in Large-scale Recommendation

    Rongcheng Lin, Yan Sun, Jamey Zhang +4

    cs.IRcs.AIarXiv:2608.16797v12026
  56. Skill2Query: Exploiting Skill Structure to Generate Pseudo-Queries for Agent Skill Retrieval

    Lihui Ding, Zihan Guo, Bingwei Lu +5

    cs.CLcs.IRarXiv:2608.16071v12026
  57. A Theoretical Framework for Parallel Lifelong MAPF Using Group Decentralized Planning

    Alex DeWeese, Jiaoyang Li, Guannan Qu

    cs.MAcs.AIcs.ROarXiv:2608.17928v12026
  58. Mixture-of-Expert Blocks Contain Strong Hallucination Detection Signals

    Joao Fonseca, Rodrigo Rodrigues, Paolo Romano

    cs.AIcs.LGarXiv:2608.17687v12026
  59. Analysis of Types of Inquiries in Student-AI Interaction: A case study of two CS2 tasks

    Matin Amoozadeh, Amin Alipour

    cs.HCcs.AIarXiv:2608.17919v12026
  60. Learnware for CSI Feedback: Scene-specific Small Models Can Do Big

    Xiangyi Li, Jiajia Guo, Chao-Kai Wen +3

    cs.ITcs.AIeess.SParXiv:2608.17760v12026