Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

18,781 to 18,840 of 61,351

  1. HoliTom: Holistic Token Merging for Fast Video Large Language Models

    Kele Shao, Keda Tao, Can Qin +3

    cs.CVarXiv:2505.21334v32025
  2. PetIGA: A Framework for High-Performance Isogeometric Analysis

    Lisandro Dalcin, Nathan Collier, Philippe Vignal +2

    cs.MSmath.NAarXiv:1305.4452v32013
  3. PixNerd: Pixel Neural Field Diffusion

    Shuai Wang, Ziteng Gao, Chenhui Zhu +2

    cs.CVarXiv:2507.23268v22025
  4. Federated Neural Collaborative Filtering

    Vasileios Perifanis, Pavlos S. Efraimidis

    cs.IRcs.CRcs.LGarXiv:2106.04405v22021
  5. O-CNN: Octree-based Convolutional Neural Networks for 3D Shape Analysis

    Peng-Shuai Wang, Yang Liu, Yu-Xiao Guo +2

    cs.CVarXiv:1712.01537v12017
  6. Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling

    Haoyu Wu, Diankun Wu, Tianyu He +4

    cs.CVcs.AIarXiv:2507.07982v22025
  7. Spatially Aware World Action Model via Geometric Latent Diffusion

    Javier Alejandro Lopetegui Gonzalez, Paul Pacaud, Cordelia Schmid

    cs.CVcs.ROarXiv:2609.02531v12026
  8. Canary-1B-v2 & Parakeet-TDT-0.6B-v3: Efficient and High-Performance Models for Multilingual ASR and AST

    Monica Sekoyan, Nithin Rao Koluguri, Nune Tadevosyan +5

    cs.CLeess.ASarXiv:2509.14128v22025
  9. Soft robotic suits: State of the art, core technologies and open challenges

    Michele Xiloyannis, Ryan Alicea, Anna-Maria Georgarakis +4

    cs.ROeess.SYarXiv:2105.10588v22021
  10. VLAW: Iterative Co-Improvement of Vision-Language-Action Policy and World Model

    Yanjiang Guo, Tony Lee, Lucy Xiaoyang Shi +3

    cs.ROarXiv:2602.12063v22026
  11. Quantum Workload Privacy Beyond Data Confidentiality

    Shaunak Suresh Pawar, Samuel Punch, Krishnendu Guha

    cs.ETarXiv:2609.02323v12026
  12. Knowledge Graph Embedding: A Survey from the Perspective of Representation Spaces

    Jiahang Cao, Jinyuan Fang, Zaiqiao Meng +1

    cs.LGcs.AIcs.CLarXiv:2211.03536v22022
  13. MST: Masked Self-Supervised Transformer for Visual Representation

    Zhaowen Li, Zhiyang Chen, Fan Yang +8

    cs.CVarXiv:2106.05656v22021
  14. SciToolAgent: A Knowledge Graph-Driven Scientific Agent for Multi-Tool Integration

    Keyan Ding, Jing Yu, Junjie Huang +3

    cs.AIcs.CLarXiv:2507.20280v12025
  15. On the Adversarial Robustness of Multi-Modal Foundation Models

    Christian Schlarmann, Matthias Hein

    cs.LGcs.AIcs.CRarXiv:2308.10741v12023
  16. AceSpec: An Asymmetric Edge-Cloud Collaborative Framework for Communication-Efficient LLM Inference

    Yida Zhang, Zhiyong Gao, Shuaibing Yue +2

    cs.DCarXiv:2609.02514v12026
  17. Reward Shaping to Mitigate Reward Hacking in RLHF

    Jiayi Fu, Xuandong Zhao, Chengyuan Yao +2

    cs.LGcs.AIcs.CLarXiv:2502.18770v72025
  18. When Do Redundant Requests Reduce Latency ?

    Nihar B. Shah, Kangwook Lee, Kannan Ramchandran

    cs.NIcs.DCcs.PFarXiv:1311.2851v12013
  19. Helly-Type Theorems for Splitting Point Sets

    Lidor Portal, Natan Rubin

    math.COcs.CGcs.DMarXiv:2609.02180v22026
  20. 3D Face Morphable Models "In-the-Wild"

    James Booth, Epameinondas Antonakos, Stylianos Ploumpis +3

    cs.CVarXiv:1701.05360v12017
  21. Image-Grounded Conversations: Multimodal Context for Natural Question and Response Generation

    Nasrin Mostafazadeh, Chris Brockett, Bill Dolan +4

    cs.CLcs.AIcs.CVarXiv:1701.08251v22017
  22. BRISC: Annotated Dataset for Brain Tumor Segmentation and Classification

    Amirreza Fateh, Yasin Rezvani, Sara Moayedi +4

    eess.IVcs.CVarXiv:2506.14318v52025
  23. Systematic Inequalities in Language Technology Performance across the World's Languages

    Damián Blasi, Antonios Anastasopoulos, Graham Neubig

    cs.CLarXiv:2110.06733v12021
  24. BuildOcc: A Large Language Model Occupant Agent Platform for Building Energy Research

    Wooyoung Jung

    cs.HCarXiv:2609.02729v12026
  25. CAFE: Catastrophic Data Leakage in Vertical Federated Learning

    Xiao Jin, Pin-Yu Chen, Chia-Yi Hsu +2

    cs.LGcs.AIarXiv:2110.15122v42021
  26. Reinforcement Learning Optimization for Large-Scale Learning: An Efficient and User-Friendly Scaling Library

    Weixun Wang, Shaopan Xiong, Gengru Chen +38

    cs.LGcs.DCarXiv:2506.06122v12025
  27. When Does Authorization End? Effect Closure at Provider Boundaries

    Igor Santos-Grueiro

    cs.CRcs.DCarXiv:2609.02866v12026
  28. Pico-Banana-400K: A Large-Scale Dataset for Text-Guided Image Editing

    Yusu Qian, Eli Bocek-Rivele, Liangchen Song +5

    cs.CVcs.CLcs.LGarXiv:2510.19808v12025
  29. A Survey on Over-the-Air Computation

    Alphan Sahin, Rui Yang

    cs.ITcs.AIeess.SParXiv:2210.11350v52022
  30. Compressed Chain of Thought: Efficient Reasoning Through Dense Representations

    Jeffrey Cheng, Benjamin Van Durme

    cs.CLarXiv:2412.13171v12024
  31. Expanding the Capabilities of Reinforcement Learning via Text Feedback

    Yuda Song, Lili Chen, Fahim Tajwar +5

    cs.LGarXiv:2602.02482v22026
  32. Traction force microscopy on soft elastic substrates: a guide to recent computational advances

    Ulrich S. Schwarz, Jerome R. D. Soine

    q-bio.QMcond-mat.softq-bio.CBarXiv:1506.02394v12015
  33. Evaluating ML-based Intrusion Detection Systems: The Illusion of Model Efficacy

    Achilleas Spanos, Ioanna Kantzavelou

    cs.CRarXiv:2609.02469v12026
  34. moco: Fast Motion Correction for Calcium Imaging

    Alexander Dubbs, James Guevara, Darcy S. Peterka +1

    cs.CVarXiv:1506.06039v12015
  35. Age of Information: The Gamma Awakening

    Elie Najm, Rajai Nasser

    cs.ITarXiv:1604.01286v12016
  36. Attention Strategies for Multi-Source Sequence-to-Sequence Learning

    Jindřich Libovický, Jindřich Helcl

    cs.CLcs.NEarXiv:1704.06567v12017
  37. Do Better Imagined Rollouts Mean Better Robot Control? A Controlled Study of World-Model Evaluation Under Feedback

    Dharini Raghavan, Amritpal Singh

    cs.ROarXiv:2609.02811v12026
  38. An unscented Kalman filter method for real time input-parameter-state estimation

    Marios Impraimakis, Andrew W. Smyth

    eess.SPcs.AIcs.CVarXiv:2511.02717v12025
  39. Unrolled Optimization with Deep Priors

    Steven Diamond, Vincent Sitzmann, Felix Heide +1

    cs.CVarXiv:1705.08041v22017
  40. Flag fault-tolerant error correction with arbitrary distance codes

    Christopher Chamberland, Michael E. Beverland

    quant-pharXiv:1708.02246v32017
  41. From Proxy Learning to Driving Decisions: A Transfer-Based Framework for Evaluating Future-Aware Autonomous Driving Planners

    Yikai Wu

    cs.ROarXiv:2609.02688v12026
  42. Large-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image Understanding

    Zhongyi Shui, Jianpeng Zhang, Weiwei Cao +8

    cs.CVarXiv:2501.14548v12025
  43. FPGA/DNN Co-Design: An Efficient Design Methodology for IoT Intelligence on the Edge

    Cong Hao, Xiaofan Zhang, Yuhong Li +5

    cs.CVarXiv:1904.04421v12019
  44. Depth-Based 3D Hand Pose Estimation: From Current Achievements to Future Goals

    Shanxin Yuan, Guillermo Garcia-Hernando, Bjorn Stenger +21

    cs.CVarXiv:1712.03917v22017
  45. Hierarchy-of-Groups Policy Optimization for Long-Horizon Agentic Tasks

    Shuo He, Lang Feng, Qi Wei +3

    cs.LGcs.AIarXiv:2602.22817v12026
  46. Factual Error Correction for Abstractive Summarization Models

    Meng Cao, Yue Dong, Jiapeng Wu +1

    cs.CLcs.AIarXiv:2010.08712v22020
  47. Characterizing Text Branch Sensitivity in Medical Vision-Language Segmentation via Evidence Decoupling

    Ziquan Liu, Zhewei Zhu, Xuyang Shi

    cs.CVarXiv:2609.02663v12026
  48. Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond

    Chongyu Fan, Jinghan Jia, Yihua Zhang +3

    cs.LGcs.CLarXiv:2502.05374v42025
  49. DriveAdapter: Breaking the Coupling Barrier of Perception and Planning in End-to-End Autonomous Driving

    Xiaosong Jia, Yulu Gao, Li Chen +3

    cs.ROcs.CVarXiv:2308.00398v22023
  50. Geo-knowledge-guided GPT models improve the extraction of location descriptions from disaster-related social media messages

    Yingjie Hu, Gengchen Mai, Chris Cundy +6

    cs.CYarXiv:2310.09340v12023
  51. Information Transmission using the Nonlinear Fourier Transform, Part II: Numerical Methods

    Mansoor I. Yousefi, Frank R. Kschischang

    cs.ITarXiv:1204.0830v22012
  52. Process vs. Outcome Reward: Which is Better for Agentic RAG Reinforcement Learning

    Wenlin Zhang, Xiangyang Li, Kuicai Dong +9

    cs.IRarXiv:2505.14069v32025
  53. LaST-SR: Laplace-Inspired Steady-Transient Complex-Frequency Decomposition for Single Image Super-Resolution

    Linhao Li, Zhaojie Pan, Langkun Chen

    cs.CVarXiv:2609.02063v12026
  54. Attending to Characters in Neural Sequence Labeling Models

    Marek Rei, Gamal K. O. Crichton, Sampo Pyysalo

    cs.CLcs.LGcs.NEarXiv:1611.04361v12016
  55. Uni-Sign: Toward Unified Sign Language Understanding at Scale

    Zecheng Li, Wengang Zhou, Weichao Zhao +3

    cs.CVarXiv:2501.15187v32025
  56. Towards Causal VQA: Revealing and Reducing Spurious Correlations by Invariant and Covariant Semantic Editing

    Vedika Agarwal, Rakshith Shetty, Mario Fritz

    cs.CVcs.CLcs.LGarXiv:1912.07538v32019
  57. Differentiable Ranks and Sorting using Optimal Transport

    Marco Cuturi, Olivier Teboul, Jean-Philippe Vert

    cs.LGstat.MLarXiv:1905.11885v22019
  58. Fathom: Reference Workloads for Modern Deep Learning Methods

    Robert Adolf, Saketh Rama, Brandon Reagen +2

    cs.LGarXiv:1608.06581v12016
  59. Position-Based Quantum Cryptography: Impossibility and Constructions

    Harry Buhrman, Nishanth Chandran, Serge Fehr +4

    quant-phcs.CRarXiv:1009.2490v42010
  60. MedHELM: Holistic Evaluation of Large Language Models for Medical Tasks

    Suhana Bedi, Hejie Cui, Miguel Fuentes +78

    cs.CLcs.AIarXiv:2505.23802v22025