Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

60,961 to 61,020 of 61,216

  1. CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG

    Gyuwan Kim, Cheoneum Park, Tao Yang

    cs.CLcs.AIcs.IRarXiv:2608.07458v12026
  2. VoiceDesigner: Text-to-Voice Generation and Editing via Unified Diffusion Modeling and Data Augmentation

    Jiarui Hai, Karan Thakkar, Ke Chen +5

    eess.AScs.LGarXiv:2608.13613v12026
  3. What to Preserve, Where to Adapt: A Depth-Wise Analysis of Forgetting in Continual Gynecological Image Segmentation

    Amal Saqib, Tausifa Jan Saleem, Numan Saeed +1

    cs.CVcs.LGarXiv:2608.13660v12026
  4. UniWorld-Design: From Pixel Generation to Layer-Native Design

    Zongjian Li, Zhiyuan Yan, Chenxu Bai +9

    cs.CVarXiv:2608.03971v12026
  5. Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models

    Siming Fu, Haojun Xu, Ruizhe He +9

    cs.CVarXiv:2608.04349v12026
  6. LLaDA MoE v2: Scaling Mixture-of-Experts Diffusion Language Models

    Fengqi Zhu, Shaoxuan Xu, Jingyang Ou +11

    cs.AIarXiv:2608.03457v12026
  7. Invisible Shortcuts: Why Vision Encoders Know Your Camera

    Vladan Stojnić, Ryan Ramos, Giorgos Kordopatis-Zilos +2

    cs.CVcs.LGarXiv:2608.05424v12026
  8. Characterizing the Quality Profile of AI-Generated C++ in Production

    Michael Tran, Fred Lewis, Kun Yang +5

    cs.SEcs.AIarXiv:2608.06640v12026
  9. VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation

    Kangning Zhang, Yixing Li, Shuai Shao +9

    cs.CVcs.CLarXiv:2607.28590v12026
  10. ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities?

    Tianyi Guan, Yiding Wang, Haotong Yang +5

    cs.AIcs.CLcs.LGarXiv:2608.03874v12026
  11. K-EXAONE 2.0 Technical Report

    Eunbi Choi, Kibong Choi, Sehyun Chun +74

    cs.CLarXiv:2608.04505v12026
  12. Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT

    Pu Zhao, Changdi Yang, Yixiao Chen +4

    cs.SEcs.AIcs.ETarXiv:2608.13681v12026
  13. GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning

    Zhaoxin Yu, Qi Shen, Hengli Li +4

    cs.LGcs.CLarXiv:2608.02585v22026
  14. Reaction-Transformation-Aware Flow Matching for Generalizable Transition State Generation

    Kaipeng Zeng, Wenxi Zhai, Shengrui Xu +5

    physics.chem-phcs.AIarXiv:2608.14076v12026
  15. Motion Beyond Morphology: Bootstrapping Cross-Category Motion Transfer from Abstract Motion Representations

    Zhixue Fang, Zhimin Zhang, Bi'an Du +6

    cs.CVarXiv:2608.01628v22026
  16. CutClean: Neural Network Pruning for Privacy-Preserving Inference

    Leonardo Magliolo, Vito Paolo Pastore, Giuseppe Valenzise +1

    cs.LGcs.AIarXiv:2608.13773v12026
  17. Intelligent Detection of Mechanical, Electrical, and Plumbing (MEP) Metrics Based on 2D Floor Plans

    Tarandeep Singh Mandhiratta, ANK Zaman, Abdul-Rahman Mawlood-Yunis

    cs.CVcs.AIcs.HCarXiv:2608.14317v12026
  18. Local and Global Regimes of Geometric Complexity in Language Model Representations

    Arwa Osman, Marco Baroni, Iuri Macocco

    cs.CLarXiv:2608.14361v12026
  19. Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages

    Chih-Hsuan Yang, Anjir Ahmed Chowdhury, Cheng-Hau Yang +7

    cs.AIcs.CLcs.LGarXiv:2608.14375v12026
  20. RecipeNet: A Hierarchical Transformer for Recipe Data

    Pin-Yen Huang, Sachin Chhabra, Prasanth Sai Gouripeddi +2

    cs.LGcs.AIarXiv:2608.14505v12026
  21. The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

    Zhiheng Wang, Bo Peng, Lai Wei +1

    cs.AIarXiv:2608.06270v12026
  22. Architecture and Affordances of PLAUD: Performative Latents and Unsupervised DDSP

    Błażej Kotowski, Frederic Font

    cs.SDcs.HCcs.LGarXiv:2608.13724v12026
  23. ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning

    Jinhe Bi, Chennan Zhou, Zengjie Jin +10

    cs.AIarXiv:2608.03972v12026
  24. Convex losses and their applications to SVM, SVR, and Shallow Neural Networks

    Filippo Portera

    cs.LGarXiv:2608.14288v12026
  25. DCAS: Decoupling CLI Agent Scaffolding to Internalize Planning across Scaffolds

    Kishanthan Thangarajah, Boyuan Chen, Ahmed E. Hassan

    cs.SEarXiv:2608.06113v12026
  26. AdsWorldEngine: A Self-Evolving Conversational Advertising Agent through Orchestrator and Tool Coevolution

    Simiao Zuo, Chenhui Xu, Yimeng Jia +3

    cs.IRcs.AIarXiv:2608.13833v12026
  27. CAPEval: A Decoupled Caption Evaluation across Understanding and Generation

    Zhipeng Liu, Haochen Wang, Zhaoxiang Zhang

    cs.CVarXiv:2608.02589v12026
  28. Building AI-Intensive Software with AI: Early Results and a Cautionary Tale on Measuring Development Cost

    Victor Barros de Miranda Neves, Kiev Santos da Gama, Vinicius Cardoso Garcia

    cs.SEcs.AIcs.LGarXiv:2608.13730v12026
  29. iFAN: Inference-Aware Learning for Plain Mask Transformers

    Fang Li, Yu He, Haoyang Tong +7

    cs.CVarXiv:2608.03216v22026
  30. ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment

    Yijun Lu, Rui Ye, Jiajun Wang +4

    cs.AIarXiv:2608.05102v12026
  31. EffectLearner: World-Aware Object-Effect Reasoning for Real-World Video Object Removal

    Feier Wu, Wanke Xia, Xu He +8

    cs.CVarXiv:2608.05565v12026
  32. CADENA: Stepwise CAD Reverse Engineering

    Soslan Kabisov, Gennadiy Savrasov, Maksim Elistratov +9

    cs.CVarXiv:2608.00799v12026
  33. Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression

    Zhengpei Hu, Kai Li, Dapeng Fu +5

    cs.CLcs.LGarXiv:2608.04569v12026
  34. The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows

    Junbo Li, Boyi Liu, Canwen Xu +5

    cs.AIarXiv:2608.06714v12026
  35. Unknown Unknowns: Model Misspecification in Machine Learning for Physics

    Juan Cruz-Martinez, Carolina Cuesta-Lazaro, Alexander Held +1

    physics.data-anastro-ph.COastro-ph.GAarXiv:2608.13633v12026
  36. Agent Against Agent: An Agentic System for Automatic Prompt Injection Red Teaming

    Yanting Wang, Chenlong Yin, Runpeng Geng +1

    cs.CRarXiv:2608.05108v12026
  37. Self-Evolving Coding Agents

    Hao Zhou, Haichuan Hu, Ye Shang +1

    cs.SEarXiv:2608.03392v12026
  38. Learning Unsteady Aneurysm Hemodynamics with Physics-Informed DeepONets

    Oscar L. Cruz-Gonzalez, Valérie Deplano, Badih Ghattas

    stat.MLcs.LGphysics.flu-dynarXiv:2608.13629v12026
  39. AgilePE: Autonomous UAV Pursuit-Evasion via Self-Play Reinforcement Learning

    Wenhao Tang, Tianyang Chen, Zhejun Cui +9

    cs.ROcs.LGarXiv:2608.14135v12026
  40. Evolve Vision-Language-Action Model into an Agent with On-the-fly Tool-use

    Yi Ding, Yanzhao Yu, Xili Dai +5

    cs.ROcs.AIcs.CVarXiv:2608.14047v12026
  41. GBU-Palm: A Multimodal Video Dataset and Benchmark for Palm Presentation Attack Detection

    Yingjie Ma, Zitong Yu, Wei Jia +2

    cs.CVcs.AIarXiv:2608.14389v12026
  42. OPD-V: Visual On-Policy Self-Distillation with Modality Balance

    Aniri, Jinhe Bi, Peng Liao +5

    cs.CVcs.AIarXiv:2608.05131v22026
  43. MiniWorld: Democratizing the Training of Video World Models from Scratch

    Yian Zhao, Ruochong Zheng, Hongcan Guo +3

    cs.CVarXiv:2608.01127v22026
  44. DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation

    Junfeng Li, Junjie He, Zhide Zhong +12

    cs.ROarXiv:2608.06374v12026
  45. Skaling: Chinchilla's Exponents Meet Kaplan's Coupling

    Mathurin Videau, Badr Youbi-Idrissi, David Lopez-Paz +1

    cs.CLarXiv:2608.07222v12026
  46. VoiceChat-TTS: A Low-Latency Continuous Speech Synthesis Model for Interactive Agents

    Edresson Casanova, Jaehyeon Kim, Mariana Graterol Fuenmayor +17

    eess.AScs.CLarXiv:2608.13831v12026
  47. Interpretable MEG Decoding of Perceived Speech: Cortical Sources and the Stimulus Features That Drive Retrieval

    Ilia Semenkov, Daria Kleeva, Ivan Dakhtin +2

    cs.LGcs.SDq-bio.NCarXiv:2608.01481v12026
  48. Ego-OSCAR: Egocentric Open source Stereo CAptuRe System

    Gunjan Paul, Senthil Palanisamy, Satpal Singh Rathore +3

    cs.CVcs.ARcs.ROarXiv:2608.08285v22026
  49. On the Brittleness of Maximum Likelihood Estimation for Gaussian Process Hyperparameter Optimization

    Tyler R. Johnson, Kian Ben-Jacob, Christopher P. Muller +1

    stat.MLcs.LGstat.MEarXiv:2608.13793v12026
  50. MameLoshnLM: Yiddish Language Model and Evaluation Benchmark

    Uri Katz, Omer Goldman, Tomasz Limisiewicz +2

    cs.CLcs.AIarXiv:2608.05850v12026
  51. SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks

    Yu Zhang, Ruiqi Li, Changhao Pan +3

    eess.AScs.SDarXiv:2608.02023v22026
  52. Continual Learning in Transition

    Zhiyan Hou, Dan Zhang, Tao Feng +11

    cs.LGcs.AIarXiv:2608.06216v22026
  53. NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap

    Dasol Choi, Joonyong Park, Daegon Yu +3

    cs.CLarXiv:2608.04397v12026
  54. Mitigating Gender Bias in English to Romanian Machine Translation

    Ioana Grigore, Sergiu Nisioi

    cs.CLcs.AIarXiv:2608.08606v12026
  55. TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning

    Changle Qu, Sunhao Dai, Hengyi Cai +4

    cs.CLcs.AIarXiv:2608.04007v12026
  56. An AI4AI Framework for Visual Token Pruning

    Zhen Liu, Wenli Huang, Wei Song +3

    cs.LGcs.CVarXiv:2608.07193v12026
  57. Expected Free Energy-based Informative Path Planning for Robotic Mars Exploration

    Ajith Anil Meera, Pablo Lanillos, Wouter Kouw

    cs.ROcs.ITcs.LGarXiv:2608.14466v12026
  58. Non-Parametric Spatiotemporal Trajectory Prediction via State-Conditioned Transition Sampling

    Michael Fore, Akshay Jain, Justin Downes +2

    cs.LGarXiv:2608.14349v12026
  59. Deep Reinforcement Learning solution for pickup and delivery routing problems with time window and capacity constraints

    Andrew Soroka, Alex Meshcheryakov, Sergey Gerasimov

    cs.LGarXiv:2608.14156v12026
  60. Enfold: Folding World Model Imagination into Predictive Representations for Ultra-Efficient Embodied Control

    Weili Zeng, Yitong Xing, Fulong Liu +10

    cs.ROcs.CVarXiv:2607.26657v32026