Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,561 to 1,620 of 61,178

  1. ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals

    Bowen Qin, Yi Xie, Yesheng Liu +1

    cs.LGcs.CLarXiv:2609.16816v12026
  2. SkillMOO: Multi-Objective Optimization of Agent Skills for Software Engineering

    Jingzhi Gong, Ruizhen Gu, Zhiwei Fei +7

    cs.SEcs.AIarXiv:2604.09297v32026
  3. Motion Mamba: Efficient and Long Sequence Motion Generation

    Zeyu Zhang, Akide Liu, Ian Reid +3

    cs.CVarXiv:2403.07487v42024
  4. PTB-TIR: A Thermal Infrared Pedestrian Tracking Benchmark

    Qiao Liu, Zhenyu He, Xin Li +1

    cs.CVarXiv:1801.05944v32018
  5. Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery

    Wenhao Li, Xiu Su, Dan Niu +6

    cs.ROarXiv:2605.01191v22026
  6. Improving Coherence and Consistency in Neural Sequence Models with Dual-System, Neuro-Symbolic Reasoning

    Maxwell Nye, Michael Henry Tessler, Joshua B. Tenenbaum +1

    cs.AIcs.CLcs.LGarXiv:2107.02794v22021
  7. FTCircuitBench: A Benchmark Suite for Fault-Tolerant Quantum Compilation and Architecture

    Adrian Harkness, Shuwen Kan, Chenxu Liu +10

    quant-pharXiv:2601.03185v22026
  8. Jointly Cross- and Self-Modal Graph Attention Network for Query-Based Moment Localization

    Daizong Liu, Xiaoye Qu, Xiao-Yang Liu +3

    cs.CVcs.IRarXiv:2008.01403v22020
  9. FLAT: Resampling Image and Text into 1D Flexible-Length Aligned Transmodal Tokens for Retrieval and Generation

    Guangyu Sun, Shlok Kumar Mishra, Wentao Bao +8

    cs.CVarXiv:2609.16591v12026
  10. SoulX-Duplug: Plug-and-Play Streaming State Prediction Module for Realtime Full-Duplex Speech Conversation

    Ruiqi Yan, Wenxi Chen, Zhanxun Liu +17

    eess.ASarXiv:2603.14877v12026
  11. "Death" of a Chatbot: Investigating and Designing Toward Psychologically Safe Endings for Human-AI Relationships

    Rachel Poonsiriwong, Chayapatr Archiwaranguprok, Pat Pataranutaporn

    cs.HCcs.AIarXiv:2602.07193v22026
  12. GraspNeRF: Multiview-based 6-DoF Grasp Detection for Transparent and Specular Objects Using Generalizable NeRF

    Qiyu Dai, Yan Zhu, Yiran Geng +3

    cs.ROcs.CVarXiv:2210.06575v32022
  13. EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards

    Ruixiang Wang, Qingming Liu, Yueci Deng +3

    cs.ROcs.AIarXiv:2603.17808v22026
  14. DreamerAD: Efficient Reinforcement Learning via Latent World Model for Autonomous Driving

    Pengxuan Yang, Yupeng Zheng, Deheng Qian +11

    cs.LGcs.ROarXiv:2603.24587v22026
  15. Automatic Curriculum Generation for Learning Adaptation in Networking

    Zhengxu Xia, Yajie Zhou, Francis Y. Yan +1

    cs.NIarXiv:2202.05940v22022
  16. Emergence World: Adversarial Stress-Testing of Long-Horizon Multi-Agent Systems

    Deepak Akkil, Tamer Abuelsaad, Karthik Vikram +5

    cs.MAarXiv:2609.17320v12026
  17. Pick Your Poison: Learning to Select Poison Sets for Stronger LLM Backdoor Attacks

    Aashiq Muhamed, Mona T. Diab, Virginia Smith +2

    cs.LGcs.AIcs.CLarXiv:2609.15029v12026
  18. Text2CAD-Bench: A Benchmark for LLM-based Text-to-Parametric CAD Generation

    Liang Wang, Heng Meng, Zekai Xiang +4

    cs.LGarXiv:2605.18430v12026
  19. PyThaiNLP: Thai Natural Language Processing in Python

    Wannaphong Phatthiyaphaibun, Korakot Chaovavanich, Charin Polpanumas +6

    cs.CLarXiv:2312.04649v12023
  20. When Agents Slow Down: Understanding LLM Agents' Test-Time Strategies via Elo-per-token Analysis

    Kaiyuan Liu, Qiuyang Mang, Bo Peng +6

    cs.CLarXiv:2609.15309v12026
  21. AlayaVista: Streaming World Modeling from Panoramic States to Perspective Video

    Jiaming Tan, Mingliang Zhai, Zhen Li +3

    cs.CVarXiv:2609.14462v12026
  22. Quantum-Selected Configuration Interaction: classical diagonalization of Hamiltonians in subspaces selected by quantum computers

    Keita Kanno, Masaya Kohda, Ryosuke Imai +4

    quant-pharXiv:2302.11320v12023
  23. AgentIR: Reasoning-Aware Retrieval for Deep Research Agents

    Zijian Chen, Xueguang Ma, Shengyao Zhuang +3

    cs.CLarXiv:2603.04384v32026
  24. MemoRAG: Boosting Long Context Processing with Global Memory-Enhanced Retrieval Augmentation

    Hongjin Qian, Zheng Liu, Peitian Zhang +4

    cs.CLcs.AIarXiv:2409.05591v32024
  25. ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement

    Siwei Wu, Jincheng Ren, Yizhi Li +11

    cs.CLarXiv:2609.14857v12026
  26. FP8 Quantization: The Power of the Exponent

    Andrey Kuzmin, Mart Van Baalen, Yuwei Ren +3

    cs.LGarXiv:2208.09225v22022
  27. Power-Measurement-Based Channel Autocorrelation Estimation for IRS-Assisted Wideband Communications

    He Sun, Lipeng Zhu, Weidong Mei +1

    cs.ITarXiv:2502.11346v12025
  28. Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation

    Niantong Li, Guangzheng Hu, Weixu Qiao +35

    cs.CVarXiv:2605.28091v22026
  29. Discovery Foundation Models: Toward Open-Ended Discovery Intelligence

    Ling Yang, Zhenfei Yin, Yingcheng Wu

    cs.CLarXiv:2609.15973v12026
  30. Mind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States

    Zixuan Wang, Yufan Zhou, Jinzhou Tang +16

    cs.CLcs.LGarXiv:2609.15972v12026
  31. How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus

    Ilya Koziev, Leonid Sinev, Ivan Oseledets

    cs.CLcs.AIarXiv:2609.15504v12026
  32. Beyond Quacking: Deep Integration of Language Models and RAG into DuckDB

    Anas Dorbani, Sunny Yasser, Jimmy Lin +1

    cs.DBcs.AIcs.IRarXiv:2504.01157v12025
  33. HazardAuditor: From Executable Threats to Safer Computer-Use Agents

    Yunhao Feng, Ruixiao Lin, Ming Wen +5

    cs.AIarXiv:2609.15134v12026
  34. Frequency Perception Network for Camouflaged Object Detection

    Runmin Cong, Mengyao Sun, Sanyi Zhang +3

    cs.CVarXiv:2308.08924v22023
  35. CORE: Simple and Effective Session-based Recommendation within Consistent Representation Space

    Yupeng Hou, Binbin Hu, Zhiqiang Zhang +1

    cs.IRcs.AIarXiv:2204.11067v12022
  36. Towards Zero-Shot Scale-Aware Monocular Depth Estimation

    Vitor Guizilini, Igor Vasiljevic, Dian Chen +2

    cs.CVcs.LGarXiv:2306.17253v12023
  37. AdaRubric: Task-Adaptive Rubrics for Reliable LLM Agent Evaluation and Reward Learning

    Liang Ding

    cs.AIcs.CLarXiv:2603.21362v32026
  38. PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control

    Chuhao Chen, Peter Wonka, Chaoyang Wang +4

    cs.CVcs.AIcs.GRarXiv:2609.17521v12026
  39. Robust shadow estimation

    Senrui Chen, Wenjun Yu, Pei Zeng +1

    quant-pharXiv:2011.09636v22020
  40. FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems

    Binghao Huang, Yunzhu Li

    cs.ROcs.AIcs.LGarXiv:2604.28156v12026
  41. OCNLI: Original Chinese Natural Language Inference

    Hai Hu, Kyle Richardson, Liang Xu +3

    cs.CLarXiv:2010.05444v12020
  42. BVB: Benchmarking Agentic Video Understanding via Programmatic Reconstruction in Blender

    Yolo Y. Tang, Daiki Shimada, Jiayue Meng +14

    cs.CVarXiv:2609.15478v12026
  43. Fg-T2M++: LLMs-Augmented Fine-Grained Text Driven Human Motion Generation

    Yin Wang, Mu Li, Jiapeng Liu +4

    cs.CVarXiv:2502.05534v12025
  44. Multi-Waveguide Pinching Antennas for ISAC

    Weihao Mao, Yang Lu, Yanqing Xu +3

    cs.ITeess.SParXiv:2505.24307v12025
  45. KaiNinja: Extending Native 3D Generators to the Part Level

    Ruihan Yu, Lian Fu, Muyao Niu +9

    cs.GRcs.AIcs.CVarXiv:2609.15659v22026
  46. How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions

    Ningzhi Tang, Chaoran Chen, Gelei Xu +5

    cs.SEcs.AIcs.HCarXiv:2605.29442v22026
  47. LynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows

    Xiaofeng Mao, Peijia Lin, Shaohao Rui +3

    cs.CVarXiv:2609.15863v12026
  48. Modality-Autoregressive World-Action Models

    Adam Hung, Bardienus P. Duisterhof, Deva Ramanan +1

    cs.ROarXiv:2609.17524v12026
  49. MS2: Multi-Document Summarization of Medical Studies

    Jay DeYoung, Iz Beltagy, Madeleine van Zuylen +2

    cs.CLcs.AIcs.LGarXiv:2104.06486v32021
  50. AudioDec: An Open-source Streaming High-fidelity Neural Audio Codec

    Yi-Chiao Wu, Israel D. Gebru, Dejan Marković +1

    eess.ASarXiv:2305.16608v12023
  51. Maximum Entropy Gain Exploration for Long Horizon Multi-goal Reinforcement Learning

    Silviu Pitis, Harris Chan, Stephen Zhao +2

    cs.LGcs.AIcs.ROarXiv:2007.02832v12020
  52. CXR-CLIP: Toward Large Scale Chest X-ray Language-Image Pre-training

    Kihyun You, Jawook Gu, Jiyeon Ham +5

    cs.CVcs.LGarXiv:2310.13292v12023
  53. Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation

    Yunhai Feng, Jiaming Han, Zhuoran Yang +3

    cs.ROcs.AIcs.LGarXiv:2502.16707v12025
  54. TJ4DRadSet: A 4D Radar Dataset for Autonomous Driving

    Lianqing Zheng, Zhixiong Ma, Xichan Zhu +9

    cs.CVcs.AIarXiv:2204.13483v32022
  55. RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments

    Sibo Zhu, Shicheng Fan, Xinyue Wang +3

    cs.AIcs.CLcs.CVarXiv:2609.15364v12026
  56. Integrated Sensing and Communications for Pinching-Antenna Systems (PASS)

    Zheng Zhang, Zhaolin Wang, Xidong Mu +3

    cs.ITarXiv:2504.07709v32025
  57. Dream-RSI: Recursive Self-Improvement through Evolving Worlds

    Tong Zheng, Xidong Wu, Zheng Zhang +14

    cs.CLarXiv:2609.14858v12026
  58. Labels Are Not Endpoints: Treatment Leakage and Construct Validity in MCP Agent Security Evaluation

    Rana Muhammad Ahmed, Sabahat Abbas

    cs.CRcs.AIarXiv:2608.12880v12026
  59. EventVAD: Training-Free Event-Aware Video Anomaly Detection

    Yihua Shao, Haojin He, Sijie Li +11

    cs.CVarXiv:2504.13092v32025
  60. PhysBrain 1.5: From Vision-Language Models to Physical Foundation Models

    DeepCybo Team, Yu Bin, Haipeng Cao +51

    cs.CVcs.ROarXiv:2609.14973v12026