Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
60,601 to 60,660 of 61,224
Vesta: A Generalist Embodied Reasoning Model
Johan Bjorck, Zhiqi Li, Yunze Man +29
cs.ROcs.AIarXiv:2606.20905v12026EBench: Elemental Diagnosis of Generalist Mobile Manipulation Policies
Ning Gao, Jinliang Zheng, Xing Gao +22
cs.ROarXiv:2606.18239v22026Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do
Zhuoran Jin, Kejian Zhu, Hongbang Yuan +5
cs.CLcs.AIcs.CVarXiv:2606.22565v12026IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation
Zixuan Li, Haokun Lin, Yicheng Xiao +10
cs.CVcs.AIarXiv:2606.24849v12026Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation
Julia Belikova, Rauf Parchiev, Evgeny Egorov +4
cs.AIcs.CLcs.SEarXiv:2606.23127v12026SimFoundry: Modular and Automated Scene Generation for Policy Learning and Evaluation
Nadun Ranawaka, Josiah Wong, Wei-Lin Pai +15
cs.ROarXiv:2606.28276v42026LISA: Likelihood Score Alignment for Visual-condition Controllable Generation
Yanghao Wang, Hongxu Chen, Jiazhen Liu +4
cs.CVarXiv:2606.27192v12026Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints
Fangzheng Li, Aimin Zhang, Chen Lv
cs.CLarXiv:2606.25605v12026Simplified Sparse Attention via Gist Tokens
Yuzhen Mao, Michael Y. Li, Emily B. Fox
cs.LGarXiv:2604.20920v22026TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents
Shoufa Chen, Luyuan Wang, Xuan Yang +7
cs.SEcs.AIarXiv:2606.28480v12026Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction
Chenguang Wang, Ming Li, Xinyue Zeng +4
cs.CLcs.AIcs.CYarXiv:2606.28186v32026Dockerless: Environment-Free Program Verifier for Coding Agents
Wenhao Zeng, Yuling Shi, Xiaodong Gu +10
cs.SEcs.AIarXiv:2606.28436v12026MemoBench: Benchmarking World Modeling in Dynamically Changing Environments
Haoyu Chen, Kaichen Zhou, Hang Hua +11
cs.CVarXiv:2606.27537v62026SPEAR: A Simulator for Photorealistic Embodied AI Research
Mike Roberts, Renhan Wang, Rushikesh Zawar +10
cs.CVcs.AIcs.GRarXiv:2607.06701v12026PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception
Yana Wei, Hongbo Peng, Yanlin Lai +14
cs.CVarXiv:2606.28322v22026Wan-Streamer v0.2: Higher Resolution, Same Latency
Lianghua Huang, Zhi-Fan Wu, Yupeng Shi +23
cs.CVcs.AIcs.GRarXiv:2607.04443v32026From Pixels to States: Rethinking Interactive World Models as Game Engines
Zhen Li, Zian Meng, Shuwei Shi +4
cs.CVarXiv:2607.14076v12026Building to the Test: Coding Agents Deliver What You Check, Not What You Requested
Yanuo Ma, Ben Kereopa-Yorke, Ben Schultz
cs.SEcs.AIarXiv:2606.28430v12026Measuring the Gap Between Human and LLM Research Ideas
Ziyu Chen, Yilun Zhao, Arman Cohan
cs.CLcs.AIarXiv:2607.01233v12026SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History
Zhiwei Li, Yong Hu
cs.LGarXiv:2606.08671v32026VIBE: Voice-Induced open-ended Bias Evaluation for Large Audio-Language Models via Real-World Speech
Yi-Cheng Lin, Yusuke Hirota, Sung-Feng Huang +1
eess.AScs.CLcs.SDarXiv:2604.17248v22026LLM Program Optimization via Retrieval Augmented Search
Sagnik Anupam, Alexander Shypula, Osbert Bastani
cs.LGarXiv:2501.18916v22025When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling
Yong Yi Bay, Kathleen A. Yearick
cs.LGcs.AIcs.CLarXiv:2606.28661v12026Parameter-Efficient Quantum-Inspired Fast Weight Programmers for Traffic-Matrix Forecasting
Kuo-Chung Peng, Jiun-Cheng Jiang, Chun-Hua Lin +3
quant-phcs.AIcs.LGarXiv:2606.27821v12026ENTRAP-VL: A Taxonomic Probe for Dual Contextual Entrainment in Vision-Language Models
Karan Goyal, Afreen Hossain, Debojyoti Das +1
cs.CVcs.AIcs.CLarXiv:2607.20092v12026SciForma: Structure-Faithful Generation of Scientific Diagrams
Yuxuan Luo, Peng Zhang, Xinjie Zhang +3
cs.CVcs.GRcs.LGarXiv:2607.18091v12026Recurrent Sinusoidal INRs for Efficient High-Fidelity Representation
Hyunmin Cho, Jaejun Yoo, Kyong Hwan Jin
cs.CVarXiv:2607.21485v12026TableVerse: A Large-scale Tabletop Dataset with Real-world Grounded Layouts for Generalizable Manipulation
Boyuan Wang, Yue Zhang, Xutao Xue +2
cs.ROarXiv:2607.21017v12026ATSplat: Compact Feed-forward 3D Gaussian Splatting with Adaptive Token Expansion
In Cho, Jeonghwan Cho, Mijin Yoo +2
cs.CVarXiv:2607.20417v22026Measuring Cross-Task Behavioral Consistency in Language Model Agents
Amritesh Banerjee, Pranil Raichura
cs.AIarXiv:2608.13598v12026Hard Cases, Bad Labels: Testing Error Exposure and Error Location in Uncertainty Sampling Under Bounded Label Noise
John Myron Uy
cs.LGarXiv:2608.13601v12026No Universal Signal Predicts Sample-Level LLM Regression under Version Updates
Jia Sheng, Yiwei Lu
cs.AIcs.CLcs.LGarXiv:2608.13607v12026HI-MeshGraphNets: Efficient and Accurate Mesh-based Physics Learning with Hierarchical Multi-scale Graph Neural Networks
SiHun Lee, Dong-Hyuk Park, Taesoo Bang +1
cs.LGarXiv:2608.13827v12026Trajectory Dynamics in Self-Supervised Learning Latent Space for Audio Deepfake Detection
Tomás Andrade Weber
eess.AScs.LGcs.SDarXiv:2608.13817v12026Language-Specific Gaps in AI Safety Training Datasets
Chialuka Prisca-Mary Onuoha, Bright Etornam Sunu, Rashidat Sikiru
cs.CYcs.LGarXiv:2608.13695v12026The Integer Alibi: Localizing Cross-Kernel Divergence in INT8-Quantized LLM Inference
Teng-Ruei Chen
cs.LGarXiv:2608.13756v12026Asymmetric Discourse Homogenization and Shared Language Technology: Evidence from Reddit
Fengming Liu
cs.CYcs.CLcs.SIarXiv:2608.13674v12026Optimal Power Allocation and AI Receiver Design for Superimposed DMRS and Data Transmission
Sha Hu, Zhongwang Fu
cs.ITcs.AIarXiv:2608.13809v12026Does ISO-Grounded NFR Specification Improve LLM Code Generation? A Comparison of Rich and Structured Interventions against a Natural-Language Baseline
Joào Pedro Monteiro Pereira, Vinicius Cardoso Garcia
cs.SEcs.AIcs.LGarXiv:2608.13742v12026SDO: Subspace Deconflicting Operator for Multi-Adapter Composition
Zhongsheng Wang, Zhedong Lin, Qian Liu +2
cs.AIarXiv:2608.13820v12026Explanation Multiplicity: Circuit-Level Interpretability Evidence Does Not Survive Defensible Analytic Variation
Ajay Pravin Mahale
cs.AIarXiv:2608.13754v12026Ontology-Grounded Project Memory for Coding Agents
James Adam
cs.AIcs.SEarXiv:2608.13662v12026A Calibrated Test of Internal Action Maps: State Signals Without Global Affine Closure
Dekun Yang
cs.AIcs.CLcs.LGarXiv:2608.13626v12026High-dimensional nonparametric changepoint detection via low-rank degree-two density projection
Guoqing Zhang, Zhaixin Chen
cs.LGstat.MLarXiv:2608.13922v12026SLAM in Low-Light Environments: Project Report
Oleh Basystyi, Anna Stasyshyn, Oleksandr Kosovan +1
cs.ROcs.CVarXiv:2607.17699v12026How Fast Can Reward Models Score? A Systems Study of C++ and PyTorch Inference Runtimes for RLHF
Venkata Naga Sai Vishnu Rohit Pulipaka, Anish Katta, Deva Rohit Reddy Peddireddy
cs.LGarXiv:2607.19712v12026Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges
Tuo Liang, Zhe Hu, Disheng Liu +2
cs.CLcs.AIcs.MMarXiv:2607.19011v12026Multimodal Speaker Verification as a Threat to Speaker Anonymization
Ashi Garg, Cristina Aggazzotti, Leibny Paola García-Perera +1
eess.ASarXiv:2607.19636v12026Self-State Attacks on Self-Hosted AI Agents: How Far Can OS Defenses Go?
Yimeng Chen, Nathanaël Denis, Roberto Di Pietro +1
cs.CRcs.AIcs.CLarXiv:2607.17986v12026Evidence Attribution in Visual Document Understanding without Coordinates or Region Labels
Zhuchenyang Liu, Yao Zhang, Yu Xiao
cs.CVcs.CLcs.IRarXiv:2607.24651v12026FlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal Applications
Krish Agarwal, Zhuoming Chen, Yanyuan Qin +3
cs.LGarXiv:2607.18171v22026Differentiable Logic Gate Networks for Low-Latency EEG Classification on Edge Devices
Shyamal Y. Dharia, Stephen D. Smith, Camilo E. Valderrama
cs.LGcs.AIarXiv:2607.18149v12026Token-Level Off-Policy Learning for Faithful Generation Under Distribution Shift
Zitong Huang, Gustavo Lucas Carvalho, Deqing Fu +1
cs.CLcs.LGarXiv:2607.17524v12026Closing the Loop: Training-Free Revisit Consistency for Autoregressive Generative Rendering
Wenchao Ma, Changran Liu, Sharon X. Huang +1
cs.CVarXiv:2607.21848v22026AI Tour Meeting: Group Travel Planning by LLM Agents
Daisuke Kikuta
cs.AIcs.CLcs.MAarXiv:2607.18806v12026Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations
Hiskias Dingeto
cs.AIcs.CLarXiv:2607.20379v12026Delineate Anything v2: A Global Foundation Model for Field Delineation
Mykola Lavreniuk, Nataliia Kussul, Andrii Shelestov +4
cs.CVarXiv:2607.19069v22026Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning
Md Tanvirul Alam
cs.CVarXiv:2607.19790v12026DriveDNA: A Large-Scale Multimodal Naturalistic Driving Dataset and Benchmark for Driving Style Identification
Yuhang Wang, Lingyao Li, Hao Zhou
cs.LGarXiv:2607.23822v12026EduPanel: A Three-Agent LLM Judge for Teaching Videos -- Reliability, Complementarity, and Human Trust Calibration
Jia-Kai Dong, Yi-Cheng Lin, Hung-yi Lee
cs.HCcs.AIarXiv:2607.18529v22026