Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

54,061 to 54,120 of 61,092

  1. Model Context Protocol (MCP) Tool Descriptions Are Smelly! Towards Improving AI Agent Efficiency with Augmented MCP Tool Descriptions

    Mohammed Mehedi Hasan, Hao Li, Gopi Krishnan Rajbahadur +2

    cs.SEcs.ETarXiv:2602.14878v32026
  2. Embed-RL: Reinforcement Learning for Reasoning-Driven Multimodal Embeddings

    Haonan Jiang, Yuji Wang, Yongjie Zhu +5

    cs.CVarXiv:2602.13823v32026
  3. Latent Ordinal Evidence, Misaligned Outputs: Inference-Time Ordinal Lens Alignment for Multimodal LLMs

    Haiming Li, Yingsheng Liu, Jingmin Zhu +5

    cs.CVarXiv:2608.20999v12026
  4. On human motion prediction using recurrent neural networks

    Julieta Martinez, Michael J. Black, Javier Romero

    cs.CVarXiv:1705.02445v12017
  5. Fast Coordinated Bimanual Motion Planning With Hard Constraints

    Borna Paro, Luka Petrović, Ivan Marković

    cs.ROarXiv:2608.20946v12026
  6. CSI: A Hybrid Deep Model for Fake News Detection

    Natali Ruchansky, Sungyong Seo, Yan Liu

    cs.LGcs.SIarXiv:1703.06959v42017
  7. Grounding Image Matching in 3D with MASt3R

    Vincent Leroy, Yohann Cabon, Jérôme Revaud

    cs.CVarXiv:2406.09756v12024
  8. Predicting Dynamic Embedding Trajectory in Temporal Interaction Networks

    Srijan Kumar, Xikun Zhang, Jure Leskovec

    cs.SIcs.CYcs.LGarXiv:1908.01207v12019
  9. Unified Vision-Language Pre-Training for Image Captioning and VQA

    Luowei Zhou, Hamid Palangi, Lei Zhang +3

    cs.CVarXiv:1909.11059v32019
  10. SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

    Gabriel Orlanski, Devjeet Roy, Alexander Yun +7

    cs.SEcs.AIcs.CLarXiv:2603.24755v22026
  11. Learning Representations for Automatic Colorization

    Gustav Larsson, Michael Maire, Gregory Shakhnarovich

    cs.CVarXiv:1603.06668v32016
  12. MobileBERT: a Compact Task-Agnostic BERT for Resource-Limited Devices

    Zhiqing Sun, Hongkun Yu, Xiaodan Song +3

    cs.CLcs.LGarXiv:2004.02984v22020
  13. Power of data in quantum machine learning

    Hsin-Yuan Huang, Michael Broughton, Masoud Mohseni +4

    quant-phcs.LGarXiv:2011.01938v22020
  14. Top-down Neural Attention by Excitation Backprop

    Jianming Zhang, Zhe Lin, Jonathan Brandt +2

    cs.CVarXiv:1608.00507v12016
  15. M2Depth: Unifying Monocular Depth Foundation Priors with Multi-View Stereo

    Byeonggwon Lee, Sanggi Lee, Siwoo Lee +2

    cs.CVarXiv:2608.20788v12026
  16. Tracking without bells and whistles

    Philipp Bergmann, Tim Meinhardt, Laura Leal-Taixe

    cs.CVarXiv:1903.05625v32019
  17. EmotionDialogCN: A Spontaneous Multimodal Dataset for Mandarin Emotional Dialogue

    Yi Zheng, Yifan Xu, Yan Zhou +8

    cs.CVarXiv:2608.20905v12026
  18. Generating Multi-view Adversarial Examples for Visual Geometry Grounded Transformer

    Qi Song, Ziyuan Luo, Haoliang Han +1

    cs.CVarXiv:2608.20748v12026
  19. Distilling Black-Box Machine Learning into a Small, Self-Explaining Language Model for Learning Analytics

    Chenguang Pan, Airui Meng, Youmi Suk

    cs.HCcs.CYarXiv:2608.21165v12026
  20. AgentProcessBench: Diagnosing Step-Level Process Quality in Tool-Using Agents

    Shengda Fan, Xuyan Ye, Yupeng Huo +9

    cs.AIarXiv:2603.14465v22026
  21. Automatic diagnosis of the 12-lead ECG using a deep neural network

    Antônio H. Ribeiro, Manoel Horta Ribeiro, Gabriela M. M. Paixão +9

    cs.LGeess.SPstat.MLarXiv:1904.01949v22019
  22. Gradient based sample selection for online continual learning

    Rahaf Aljundi, Min Lin, Baptiste Goujaud +1

    cs.LGcs.AIcs.CVarXiv:1903.08671v52019
  23. DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation

    Xu Guo, Fulong Ye, Qichao Sun +7

    cs.CVarXiv:2602.12160v12026
  24. Revisiting Point Cloud Classification: A New Benchmark Dataset and Classification Model on Real-World Data

    Mikaela Angelina Uy, Quang-Hieu Pham, Binh-Son Hua +2

    cs.CVarXiv:1908.04616v22019
  25. DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing

    Dianyi Wang, Ruihang Li, Feng Han +17

    cs.CVcs.AIarXiv:2602.12205v22026
  26. A Universal Part-of-Speech Tagset

    Slav Petrov, Dipanjan Das, Ryan McDonald

    cs.CLarXiv:1104.2086v12011
  27. Toward Vision Language Model-based Assessment of Clinical Quality and Usability of LGE-MR Images for Cardiac Ablation Planning

    Bipasha Kundu, Abhishek Chaturvedi, Axel W. E. Wismueller +2

    eess.IVcs.CVarXiv:2608.21180v12026
  28. JavisDiT++: Unified Modeling and Optimization for Joint Audio-Video Generation

    Kai Liu, Yanhao Zheng, Kai Wang +7

    cs.CVcs.MMcs.SDarXiv:2602.19163v12026
  29. FractalNet: Ultra-Deep Neural Networks without Residuals

    Gustav Larsson, Michael Maire, Gregory Shakhnarovich

    cs.CVarXiv:1605.07648v42016
  30. Multivariate LSTM-FCNs for Time Series Classification

    Fazle Karim, Somshubra Majumdar, Houshang Darabi +1

    cs.LGstat.MLarXiv:1801.04503v22018
  31. The VIA Annotation Software for Images, Audio and Video

    Abhishek Dutta, Andrew Zisserman

    cs.CVarXiv:1904.10699v32019
  32. LongCat-Flash-Thinking-2601 Technical Report

    Meituan LongCat Team, Anchun Gui, Bei Li +163

    cs.AIarXiv:2601.16725v22026
  33. Diffusion Probabilistic Models for 3D Point Cloud Generation

    Shitong Luo, Wei Hu

    cs.CVarXiv:2103.01458v22021
  34. GAP-SAM: A Global Artifact Prior for Generalizable AI-Generated Image Manipulation Localization

    Haozhen Yan, Siyuan Shan, Zijian Yu +4

    cs.CVarXiv:2608.20929v12026
  35. Gaia2: Benchmarking LLM Agents on Dynamic and Asynchronous Environments

    Romain Froger, Pierre Andrews, Matteo Bettini +21

    cs.AIarXiv:2602.11964v12026
  36. Bilinear Attention Networks

    Jin-Hwa Kim, Jaehyun Jun, Byoung-Tak Zhang

    cs.CVcs.AIcs.CLarXiv:1805.07932v22018
  37. MAD-GAN: Multivariate Anomaly Detection for Time Series Data with Generative Adversarial Networks

    Dan Li, Dacheng Chen, Lei Shi +3

    cs.LGstat.MLarXiv:1901.04997v12019
  38. Pneumatic Units for Logic-based Sequential Excitation (PULSE) in Wearable Haptic Devices

    Jessica Healey, Anoush Sepehri, Michael T. Tolley +1

    cs.HCcs.ROarXiv:2608.20626v12026
  39. Multicell MIMO Communications Relying on Intelligent Reflecting Surface

    Cunhua Pan, Hong Ren, Kezhi Wang +4

    eess.SParXiv:1907.10864v42019
  40. Dr. Kernel: Reinforcement Learning Done Right for Triton Kernel Generations

    Wei Liu, Jiawei Xu, Yingru Li +4

    cs.LGcs.AIcs.CLarXiv:2602.05885v22026
  41. Actional-Structural Graph Convolutional Networks for Skeleton-based Action Recognition

    Maosen Li, Siheng Chen, Xu Chen +3

    cs.CVcs.AIarXiv:1904.12659v12019
  42. Interpretation of Neural Networks is Fragile

    Amirata Ghorbani, Abubakar Abid, James Zou

    stat.MLcs.LGarXiv:1710.10547v22017
  43. SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

    Nicholas Pfaff, Thomas Cohn, Sergey Zakharov +2

    cs.ROcs.AIcs.CVarXiv:2602.09153v22026
  44. ContextBench: A Benchmark for Context Retrieval in Coding Agents

    Han Li, Letian Zhu, Bohan Zhang +7

    cs.LGarXiv:2602.05892v32026
  45. ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation

    Zedong Chu, Shichao Xie, Xiaolong Wu +41

    cs.ROcs.AIcs.CVarXiv:2602.11598v12026
  46. Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights

    Yulu Gan, Phillip Isola

    cs.LGcs.AIarXiv:2603.12228v12026
  47. ActionMesh: Animated 3D Mesh Generation with Temporal 3D Diffusion

    Remy Sabathier, David Novotny, Niloy J. Mitra +1

    cs.CVarXiv:2601.16148v22026
  48. Measurement-based quantum computation

    H. J. Briegel, D. E. Browne, W. Dür +2

    quant-pharXiv:0910.1116v22009
  49. ARISE: Agent Reasoning with Intrinsic Skill Evolution in Hierarchical Reinforcement Learning

    Yu Li, Rui Miao, Zhengling Qi +1

    cs.AIarXiv:2603.16060v22026
  50. A5-miseq: an updated pipeline to assemble microbial genomes from Illumina MiSeq data

    David Coil, Guillaume Jospin, Aaron E. Darling

    q-bio.QMq-bio.GNarXiv:1401.5130v22014
  51. K-Search: LLM Kernel Generation via Co-Evolving Intrinsic World Model

    Shiyi Cao, Ziming Mao, Joseph E. Gonzalez +1

    cs.AIarXiv:2602.19128v22026
  52. Unrolled Generative Adversarial Networks

    Luke Metz, Ben Poole, David Pfau +1

    cs.LGstat.MLarXiv:1611.02163v42016
  53. The NarrativeQA Reading Comprehension Challenge

    Tomáš Kočiský, Jonathan Schwarz, Phil Blunsom +4

    cs.CLcs.AIcs.NEarXiv:1712.07040v12017
  54. Live Artifacts: Authoring Dynamic Media via Live Layers Encapsulating Generative Specifications

    Leixian Shen, Haotian Li, Hugo Romat +3

    cs.HCarXiv:2608.20880v12026
  55. LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs

    Keda Tao, Yuhua Zheng, Jia Xu +13

    cs.CVarXiv:2603.19217v12026
  56. PhysNet: A Neural Network for Predicting Energies, Forces, Dipole Moments and Partial Charges

    Oliver T. Unke, Markus Meuwly

    physics.chem-pharXiv:1902.08408v22019
  57. FeatureBench: Benchmarking Agentic Coding for Complex Feature Development

    Qixing Zhou, Jiacheng Zhang, Haiyang Wang +9

    cs.SEcs.AIarXiv:2602.10975v12026
  58. UniXcoder: Unified Cross-Modal Pre-training for Code Representation

    Daya Guo, Shuai Lu, Nan Duan +3

    cs.CLcs.PLcs.SEarXiv:2203.03850v12022
  59. On the Time and Frequency Domain Representations of Signals for CPS Specification

    Claudio Mandrioli, Drishti Yadav, Domenico Bianculli

    cs.SEarXiv:2608.21167v12026
  60. Hacking commercial quantum cryptography systems by tailored bright illumination

    Lars Lydersen, Carlos Wiechers, Christoffer Wittmann +3

    quant-pharXiv:1008.4593v22010