Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

21,841 to 21,900 of 61,351

  1. Appearance Pointers -- Multimodal Region Control of Diffusion Transformers

    Rahul Sajnani, Yulia Gryaditskaya, Radomír Měch +2

    cs.CVcs.AIcs.GRarXiv:2607.19344v12026
  2. ExplainBench: Evaluating Code Explanations from Agents

    Zhiyuan Pan, Sungmin Kang, Imam Nur Bani Yusuf +1

    cs.SEarXiv:2607.26451v12026
  3. Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget

    Guoxuan Chen, Chufeng Xiao, Haoran Yang +30

    cs.CVcs.AIarXiv:2607.13125v22026
  4. Length Penalties Make Chain-of-Thought Less Monitorable

    Bryce Little

    cs.AIcs.CLcs.LGarXiv:2607.09786v32026
  5. AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation

    Bao Long Nguyen Huu, Atsushi Hashimoto

    cs.IRcs.AIarXiv:2607.00052v12026
  6. Confidence-Aware Tool Orchestration for Robust Video Understanding

    Yangfan He, Yujin Choi, Jaehong Yoon

    cs.CVcs.AIarXiv:2606.26904v12026
  7. Learning to Trigger: Reinforcement Learning at the Large Hadron Collider

    Zixin Ding, Shaghayegh Emami, Giovanna Salvi +7

    cs.LGcs.AIhep-exarXiv:2606.23993v32026
  8. Comparing Linear Probes with Mahalanobis Cosine Similarity

    Zhuofan Josh Ying, Peter Hase, Nikolaus Kriegeskorte

    cs.LGarXiv:2606.19603v12026
  9. Kairos: A Regret-Aware Native World-Action Model Stack for Physical AI

    Kairos Team, Fei Wang, Shan You +21

    cs.AIcs.CVarXiv:2606.16533v32026
  10. OpenThoughts-Agent: Data Recipes for Agentic Models

    Negin Raoof, Richard Zhuang, Marianna Nezhurina +47

    cs.AIarXiv:2606.24855v12026
  11. STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability

    Haipeng Luo, Qingfeng Sun, Songli Wu +4

    cs.LGcs.AIcs.CLarXiv:2606.19236v12026
  12. MeshFlow: Mesh Generation with Equivariant Flow Matching

    Qi Sun, Kiyohiro Nakayama, Jing Nathan Yan +6

    cs.GRcs.CVarXiv:2606.23489v12026
  13. MUGEN: Generating Unlearnable Graph Examples for Multiple Learning Tasks

    Ziyan Liu, Chengshuai Zhao, Huan Liu

    cs.LGarXiv:2609.00696v22026
  14. RIS-Assisted Communication Radar Coexistence: Joint Beamforming Design and Analysis

    Yinghui He, Yunlong Cai, Hao Mao +1

    cs.ITeess.SParXiv:2201.07399v12022
  15. Exploring the Political Agenda of the European Parliament Using a Dynamic Topic Modeling Approach

    Derek Greene, James P. Cross

    cs.CLcs.CYarXiv:1607.03055v12016
  16. Can Generalist Agents Automate Data Curation?

    Feiyang Kang, Hanze Li, Adam Nguyen +5

    cs.AIcs.CLcs.CVarXiv:2606.04261v12026
  17. Why Larger Models Learn More: Effects of Capacity, Interference, and Rare-Task Retention

    Jing Huang, Daniel Wurgaft, Rachit Bansal +6

    cs.LGarXiv:2605.29548v22026
  18. Token Budgets: An Empirical Catalog of 63 LLM-Agent Budget-Overrun Incidents, with an Affine-Typed Rust Mitigation as a Case Study

    Sajjad Khan

    cs.SEcs.MAcs.PLarXiv:2606.04056v12026
  19. AURA: Action-Gated Memory for Robot Policies at Constant VRAM

    Josef Chen

    cs.AIcs.ARcs.DCarXiv:2606.02775v12026
  20. Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation

    Ziyue Lin, Jiahe Hou, Hongyu Xia +6

    cs.CVarXiv:2606.01048v12026
  21. Edge AI: A Taxonomy, Systematic Review and Future Directions

    Sukhpal Singh Gill, Muhammed Golec, Jianmin Hu +12

    cs.DCarXiv:2407.04053v22024
  22. What Does an Agentic Software Engineering Benchmark Measure? Profiling Task Demands and Agent Behaviour Beyond What Category Labels Reveal

    Radin Shayanfar, Keheliya Gallaba, Ahmed E. Hassan

    cs.SEcs.CLarXiv:2609.01271v12026
  23. Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models

    Mingze Wang, Shuchen Zhu, Yuxin Fang +3

    cs.LGcs.AIstat.MLarXiv:2605.26895v12026
  24. Resource Management in Wireless Networks via Multi-Agent Deep Reinforcement Learning

    Navid Naderializadeh, Jaroslaw Sydir, Meryem Simsek +1

    cs.LGcs.ITcs.MAarXiv:2002.06215v22020
  25. Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers

    Shuhong Zheng, Michael Oechsle, Erik Sandström +3

    cs.CVcs.AIcs.GRarXiv:2605.23892v12026
  26. Reinforcement Learning to Rank in E-Commerce Search Engine: Formalization, Analysis, and Application

    Yujing Hu, Qing Da, Anxiang Zeng +2

    cs.LGarXiv:1803.00710v32018
  27. Learning Task-Specific Antibody Representations via Function-Aware Masking

    Ayan Goel, Thomas A. Walton, Amirali Aghazadeh

    cs.LGq-bio.BMarXiv:2609.00518v12026
  28. Generative Modeling with Orbit-Space Particle Flow Matching

    Sinan Wang, Jinjin He, Shenyifan Lu +3

    cs.GRcs.CVarXiv:2605.02222v12026
  29. EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

    Rong Wu, Xiaoman Wang, Jianbiao Mei +8

    cs.CLcs.AIarXiv:2510.16079v32025
  30. Micro-Defects Expose Macro-Fakes: Detecting AI-Generated Images via Local Distributional Shifts

    Boxuan Zhang, Jianing Zhu, Qifan Wang +2

    cs.CVcs.AIcs.LGarXiv:2605.09296v12026
  31. HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videos

    Jeongeun Park, Janghyeok Han, Geonung Kim +4

    cs.CVcs.GRarXiv:2605.17543v32026
  32. CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization

    Ahmed Heakl, Abdelrahman M. Shaker, Youssef Mohamed +4

    cs.LGcs.CLcs.CVarXiv:2605.19436v12026
  33. A2RBench: An Automatic Paradigm for Formally Verifiable Abstract Reasoning Benchmark Generation

    Qingchuan Ma, Yuexiao Ma, Yongkang Xie +3

    cs.AIcs.LGarXiv:2605.17278v12026
  34. Fruit Detection, Segmentation and 3D Visualisation of Environments in Apple Orchards

    Hanwen Kang, Chao Chen

    cs.CVeess.IVarXiv:1911.12889v12019
  35. Learning Multi-Level Features with Matryoshka Sparse Autoencoders

    Bart Bussmann, Noa Nabeshima, Adam Karvonen +1

    cs.LGcs.AIarXiv:2503.17547v12025
  36. $π_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

    Physical Intelligence, Bo Ai, Ali Amin +85

    cs.LGcs.ROarXiv:2604.15483v22026
  37. UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors

    Houyuan Chen, Hong Li, Xianghao Kong +8

    cs.CVarXiv:2605.00658v12026
  38. SDARE-Bench: Evaluating Large Language Models on Conversational Stigma Detection and Response in Dyadic and Group Dialogue

    Stephanie Fong, Yiwen Jiang, Zimu Wang +12

    cs.CLarXiv:2609.01548v12026
  39. Latent Preference Modeling for Cross-Session Personalized Tool Calling

    Yejin Yoon, Minseo Kim, Taeuk Kim

    cs.CLcs.AIarXiv:2604.17886v12026
  40. SUGAR: Subgraph Neural Network with Reinforcement Pooling and Self-Supervised Mutual Information Mechanism

    Qingyun Sun, Jianxin Li, Hao Peng +4

    cs.LGcs.AIarXiv:2101.08170v32021
  41. Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning

    Jiaxi Bi, Tongxu Luo, Wenyu Du +2

    cs.CLcs.LGarXiv:2604.16029v22026
  42. Trust as indicator of robot functional and social acceptance. An experimental study on user conformation to the iCub's answers

    Ilaria Gaudiello, Elisabetta Zibetti, Sebastien Lefort +2

    cs.ROcs.CYcs.HCarXiv:1510.03678v12015
  43. RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics

    Enshen Zhou, Jingkun An, Cheng Chi +8

    cs.ROcs.AIcs.CVarXiv:2506.04308v42025
  44. Semantic Richness or Geometric Reasoning? The Fragility of VLM's Visual Invariance

    Jason Qiu, Zachary Meurer, Xavier Thomas +1

    cs.CVarXiv:2604.01848v42026
  45. On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers

    Omer Dahary, Benaya Koren, Daniel Garibi +1

    cs.CVcs.AIcs.GRarXiv:2603.28762v22026
  46. ExpArt-KG: Artwork Image Description Generation through Iterative Exploration of Knowledge Graphs

    Yuta Kato, Shintaro Ozaki, Kazuki Hayashi +4

    cs.CLcs.CVarXiv:2609.00629v12026
  47. Less Gaussians, Texture More: 4K Feed-Forward Textured Splatting

    Yixing Lao, Xuyang Bai, Xiaoyang Wu +7

    cs.CVarXiv:2603.25745v12026
  48. DreamGen: Unlocking Generalization in Robot Learning through Video World Models

    Joel Jang, Seonghyeon Ye, Zongyu Lin +25

    cs.ROcs.AIcs.LGarXiv:2505.12705v22025
  49. Cognitive Mismatch in Multimodal Large Language Models for Discrete Symbol Understanding

    Yinghui Li, Jiayi Kuang, Peng Xing +11

    cs.AIcs.CVarXiv:2603.18472v22026
  50. Robust Tube-based Model Predictive Control with Koopman Operators--Extended Version

    Xinglong Zhang, Wei Pan, Riccardo Scattolini +2

    eess.SYarXiv:2108.13011v52021
  51. Live-SWE-agent: Can Software Engineering Agents Self-Evolve on the Fly?

    Chunqiu Steven Xia, Zhe Wang, Yan Yang +2

    cs.SEcs.AIcs.CLarXiv:2511.13646v32025
  52. Taking the Whys Seriously: Limitations of Counterfactual Explanations in Justification and Recourse

    Mattia Cerrato, Otto Sahlgren, Xenia Heilmann

    cs.CYcs.AIarXiv:2608.30956v12026
  53. The Deterministic Dendritic Cell Algorithm

    Julie Greensmith, Uwe Aickelin

    cs.AIcs.NEarXiv:1006.1512v12010
  54. Scaling Computer-Use Grounding via User Interface Decomposition and Synthesis

    Tianbao Xie, Jiaqi Deng, Xiaochuan Li +12

    cs.AIcs.CLcs.CVarXiv:2505.13227v32025
  55. GRADE: Benchmarking Discipline-Informed Reasoning in Image Editing

    Mingxin Liu, Ziqian Fan, Zhaokai Wang +13

    cs.CVarXiv:2603.12264v12026
  56. Time Limits in Reinforcement Learning

    Fabio Pardo, Arash Tavakoli, Vitaly Levdik +1

    cs.LGarXiv:1712.00378v42017
  57. X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents

    Salman Rahman, Liwei Jiang, James Shiffer +7

    cs.CRcs.AIcs.CLarXiv:2504.13203v22025
  58. What Makes a Good Query? Measuring the Impact of Human-Confusing Linguistic Features on LLM Performance

    William Watson, Nicole Cho, Sumitra Ganesh +1

    cs.CLcs.AIarXiv:2602.20300v12026
  59. DanceFormer: Music Conditioned 3D Dance Generation with Parametric Motion Transformer

    Buyu Li, Yongchi Zhao, Zhelun Shi +1

    cs.AIcs.CVarXiv:2103.10206v52021
  60. Memory Attention Networks for Skeleton-based Action Recognition

    Chunyu Xie, Ce Li, Baochang Zhang +4

    cs.CVarXiv:1804.08254v22018