Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

55,381 to 55,440 of 61,141

  1. SketchVLM: Vision language models can annotate images to explain thoughts and guide users

    Brandon Collins, Logan Bolton, Hung Huy Nguyen +3

    cs.CVcs.AIarXiv:2604.22875v22026
    Summaries:한국어
  2. Active RIS vs. Passive RIS: Which Will Prevail in 6G?

    Zijian Zhang, Linglong Dai, Xibi Chen +4

    cs.ITeess.SPeess.SYarXiv:2103.15154v82021
  3. Mass-Editing Memory in a Transformer

    Kevin Meng, Arnab Sen Sharma, Alex Andonian +2

    cs.CLcs.LGarXiv:2210.07229v22022
  4. Learning the solution operator of parametric partial differential equations with physics-informed DeepOnets

    Sifan Wang, Hanwen Wang, Paris Perdikaris

    cs.LGmath.NAstat.MLarXiv:2103.10974v12021
  5. From Skills to Talent: Organising Heterogeneous Agents as a Real-World Company

    Zhengxu Yu, Yu Fu, Zhiyuan He +5

    cs.AIarXiv:2604.22446v12026
  6. Knowledge Graph Convolutional Networks for Recommender Systems

    Hongwei Wang, Miao Zhao, Xing Xie +2

    cs.IRcs.LGstat.MLarXiv:1904.12575v12019
  7. Person Re-identification: Past, Present and Future

    Liang Zheng, Yi Yang, Alexander G. Hauptmann

    cs.CVarXiv:1610.02984v12016
  8. TALL: Temporal Activity Localization via Language Query

    Jiyang Gao, Chen Sun, Zhenheng Yang +1

    cs.CVarXiv:1705.02101v22017
  9. Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models

    Mohammed Safi Ur Rahman Khan, Sanjay Suryanarayanan, Tushar Anand +1

    cs.CVcs.CLarXiv:2604.21523v12026
  10. PhaseNet: A Deep-Neural-Network-Based Seismic Arrival Time Picking Method

    Weiqiang Zhu, Gregory C. Beroza

    physics.geo-phstat.AParXiv:1803.03211v12018
  11. RaV-IDP: A Reconstruction-as-Validation Framework for Faithful Intelligent Document Processing

    Pritesh Jha

    cs.CVcs.AIarXiv:2604.23644v12026
  12. Massively Multilingual Sentence Embeddings for Zero-Shot Cross-Lingual Transfer and Beyond

    Mikel Artetxe, Holger Schwenk

    cs.CLcs.AIcs.LGarXiv:1812.10464v22018
  13. Optical Quantum Computing

    Jeremy L. O'Brien

    quant-pharXiv:0803.1554v12008
  14. Reinforcement Learning for Solving the Vehicle Routing Problem

    Mohammadreza Nazari, Afshin Oroojlooy, Lawrence V. Snyder +1

    cs.AIcs.LGstat.MLarXiv:1802.04240v22018
  15. Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification

    Yunsheng Shi, Zhengjie Huang, Shikun Feng +3

    cs.LGstat.MLarXiv:2009.03509v52020
  16. XTREME: A Massively Multilingual Multi-task Benchmark for Evaluating Cross-lingual Generalization

    Junjie Hu, Sebastian Ruder, Aditya Siddhant +3

    cs.CLcs.LGarXiv:2003.11080v52020
  17. Deep Double Descent: Where Bigger Models and More Data Hurt

    Preetum Nakkiran, Gal Kaplun, Yamini Bansal +3

    cs.LGcs.CVcs.NEarXiv:1912.02292v12019
  18. Long Short-Term Memory-Networks for Machine Reading

    Jianpeng Cheng, Li Dong, Mirella Lapata

    cs.CLcs.NEarXiv:1601.06733v72016
  19. Improving Robustness of Tabular Retrieval via Representational Stability

    Kushal Raj Bhandari, Adarsh Singh, Jianxi Gao +2

    cs.CLcs.AIcs.IRarXiv:2604.24040v22026
  20. Stabilizing Efficient Reasoning with Step-Level Advantage Selection

    Han Wang, Xiaodong Yu, Jialian Wu +4

    cs.CLcs.LGarXiv:2604.24003v12026
  21. Fooling LIME and SHAP: Adversarial Attacks on Post hoc Explanation Methods

    Dylan Slack, Sophie Hilgard, Emily Jia +2

    cs.LGcs.AIstat.MLarXiv:1911.02508v22019
  22. MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills

    Yingyong Hou, Xinyuan Lao, Huimei Wang +10

    cs.AIarXiv:2604.20441v12026
  23. Efficient Agent Evaluation via Diversity-Guided User Simulation

    Itay Nakash, George Kour, Ateret Anaby-Tavor

    cs.AIarXiv:2604.21480v12026
  24. Coding for Errors and Erasures in Random Network Coding

    Ralf Koetter, Frank Kschischang

    cs.ITcs.NIarXiv:cs/0703061v22007
  25. ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks

    Mohit Shridhar, Jesse Thomason, Daniel Gordon +5

    cs.CVcs.AIcs.CLarXiv:1912.01734v22019
  26. Activation Functions in Deep Learning: A Comprehensive Survey and Benchmark

    Shiv Ram Dubey, Satish Kumar Singh, Bidyut Baran Chaudhuri

    cs.LGcs.NEarXiv:2109.14545v32021
  27. Retrieval Augmentation Reduces Hallucination in Conversation

    Kurt Shuster, Spencer Poff, Moya Chen +2

    cs.CLcs.AIarXiv:2104.07567v12021
  28. Attention Augmented Convolutional Networks

    Irwan Bello, Barret Zoph, Ashish Vaswani +2

    cs.CVarXiv:1904.09925v52019
  29. Neural 3D Mesh Renderer

    Hiroharu Kato, Yoshitaka Ushiku, Tatsuya Harada

    cs.CVcs.LGarXiv:1711.07566v12017
  30. Preferences of a Voice-First Nation: Large-Scale Pairwise Evaluation and Preference Analysis for TTS in Indian Languages

    Srija Anand, Ashwin Sankar, Ishvinder Sethi +10

    cs.CLarXiv:2604.21481v22026
  31. InternImage: Exploring Large-Scale Vision Foundation Models with Deformable Convolutions

    Wenhai Wang, Jifeng Dai, Zhe Chen +9

    cs.CVarXiv:2211.05778v42022
  32. A Caputo fractional derivative of a function with respect to another function

    Ricardo Almeida

    math.CAarXiv:1609.04775v12016
  33. Probing Visual Planning in Image Editing Models

    Zhimu Zhou, Yanpeng Zhao, Qiuyu Liao +2

    cs.CVcs.AIarXiv:2604.22868v12026
  34. Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling

    Zhen Zhang, Changyi Yang, Zijie Xia +11

    cs.CLarXiv:2604.27039v22026
  35. Discovering Agentic Safety Specifications from 1-Bit Danger Signals

    Víctor Gallego

    cs.AIcs.CLarXiv:2604.23210v12026
  36. Evaluating analytic gradients on quantum hardware

    Maria Schuld, Ville Bergholm, Christian Gogolin +2

    quant-pharXiv:1811.11184v12018
  37. A Survey on Network Embedding

    Peng Cui, Xiao Wang, Jian Pei +1

    cs.SIarXiv:1711.08752v12017
  38. IndustryAssetEQA: A Neurosymbolic Operational Intelligence System for Embodied Question Answering in Industrial Asset Maintenance

    Chathurangi Shyalika, Dhaval Patel, Amit Sheth

    cs.AIarXiv:2604.23446v12026
  39. Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence

    Yuyuan Feng, Zhishang Xiang, Chaobin Yang +32

    cs.IRcs.AIcs.ETarXiv:2608.21156v12026
  40. AnalogRetriever: Learning Cross-Modal Representations for Analog Circuit Retrieval

    Yihan Wang, Lei Li, Yao Lai +2

    cs.CVcs.AIarXiv:2604.23195v12026
  41. DeeperCut: A Deeper, Stronger, and Faster Multi-Person Pose Estimation Model

    Eldar Insafutdinov, Leonid Pishchulin, Bjoern Andres +2

    cs.CVarXiv:1605.03170v32016
  42. OceanPile: A Large-Scale Multimodal Ocean Corpus for Foundation Models

    Yida Xue, Ningyu Zhang, Tingwei Wu +5

    cs.MMcs.AIcs.CLarXiv:2605.00877v22026
  43. Data Augmentation Generative Adversarial Networks

    Antreas Antoniou, Amos Storkey, Harrison Edwards

    stat.MLcs.CVcs.LGarXiv:1711.04340v32017
  44. Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentation

    Simone Mosco, Daniel Fusaro, Alberto Pretto

    cs.CVcs.ROarXiv:2604.23604v12026
  45. Wasserstein Auto-Encoders

    Ilya Tolstikhin, Olivier Bousquet, Sylvain Gelly +1

    stat.MLcs.LGarXiv:1711.01558v42017
  46. AgentBench: Evaluating LLMs as Agents

    Xiao Liu, Hao Yu, Hanchen Zhang +19

    cs.AIcs.CLcs.LGarXiv:2308.03688v32023
  47. Quantum Kernel Advantage over Classical Collapse in Medical Foundation Model Embeddings

    Sebastian Cajas Ordóñez, Felipe Ocampo Osorio, Dax Enshan Koh +10

    quant-phcs.AIarXiv:2604.24597v12026
  48. VIBE: Video Inference for Human Body Pose and Shape Estimation

    Muhammed Kocabas, Nikos Athanasiou, Michael J. Black

    cs.CVarXiv:1912.05656v32019
  49. DenseFusion: 6D Object Pose Estimation by Iterative Dense Fusion

    Chen Wang, Danfei Xu, Yuke Zhu +4

    cs.CVcs.ROarXiv:1901.04780v12019
  50. Learning Spatio-Temporal Transformer for Visual Tracking

    Bin Yan, Houwen Peng, Jianlong Fu +2

    cs.CVarXiv:2103.17154v12021
  51. AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark

    Hongxin Li, Xiping Wang, Jingran Su +4

    cs.CVarXiv:2604.24441v12026
  52. Secure Transmission with Multiple Antennas: The MISOME Wiretap Channel

    Ashish Khisti, Gregory Wornell

    cs.ITarXiv:0708.4219v12007
  53. CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval

    Huaishao Luo, Lei Ji, Ming Zhong +4

    cs.CVarXiv:2104.08860v22021
  54. GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction

    Hongxin Li, Yuntao Chen, Zhaoxiang Zhang

    cs.CVarXiv:2604.23941v12026
  55. S^3-Rec: Self-Supervised Learning for Sequential Recommendation with Mutual Information Maximization

    Kun Zhou, Hui Wang, Wayne Xin Zhao +5

    cs.IRcs.LGarXiv:2008.07873v12020
  56. StackGAN++: Realistic Image Synthesis with Stacked Generative Adversarial Networks

    Han Zhang, Tao Xu, Hongsheng Li +4

    cs.CVcs.AIstat.MLarXiv:1710.10916v32017
  57. Composition-based Multi-Relational Graph Convolutional Networks

    Shikhar Vashishth, Soumya Sanyal, Vikram Nitin +1

    cs.LGstat.MLarXiv:1911.03082v22019
  58. ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

    Lin Chen, Jinsong Li, Xiaoyi Dong +5

    cs.CVarXiv:2311.12793v22023
  59. Enhanced LSTM for Natural Language Inference

    Qian Chen, Xiaodan Zhu, Zhenhua Ling +3

    cs.CLarXiv:1609.06038v32016
  60. BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate

    Arnon Mazza, Elad Levi

    cs.CLcs.AIcs.LGarXiv:2604.25203v12026