Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

52,081 to 52,140 of 61,132

  1. SEAR: Schema-Based Evaluation and Routing for LLM Gateways

    Zecheng Zhang, Han Zheng, Yue Xu

    cs.DBcs.AIcs.CLarXiv:2603.26728v12026
  2. OptiSight: Bridging Semantic Reasoning and Geometric Control for Embodied Navigation

    Alperen Avan, Jordi Sanchez-Riera

    cs.ROcs.CVarXiv:2608.23354v12026
  3. HigherHRNet: Scale-Aware Representation Learning for Bottom-Up Human Pose Estimation

    Bowen Cheng, Bin Xiao, Jingdong Wang +3

    cs.CVcs.LGeess.IVarXiv:1908.10357v32019
  4. Extract Free Dense Labels from CLIP

    Chong Zhou, Chen Change Loy, Bo Dai

    cs.CVcs.CLarXiv:2112.01071v22021
  5. Contrastive Representation-Guided Genetic Minority Oversampling for Imbalanced Time-Series Classification

    Wenbin Pei, Yunrong Hao, Zhen Liu +4

    cs.LGarXiv:2608.22804v12026
  6. Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior

    Antigoni-Maria Founta, Constantinos Djouvas, Despoina Chatzakou +6

    cs.SIarXiv:1802.00393v32018
  7. LagrangeGS: Non-Conservative Lagrangian System on Dynamic 3D Gaussian Splatting

    Shogo Sato, Takuhiro Kaneko, Shoichiro Takeda +4

    cs.CVarXiv:2608.22773v12026
  8. VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding

    Hu Xu, Gargi Ghosh, Po-Yao Huang +5

    cs.CVcs.CLarXiv:2109.14084v22021
  9. High Speed and High Dynamic Range Video with an Event Camera

    Henri Rebecq, René Ranftl, Vladlen Koltun +1

    cs.CVarXiv:1906.07165v12019
  10. PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

    Yanli Zhao, Andrew Gu, Rohan Varma +15

    cs.DCcs.AIcs.LGarXiv:2304.11277v22023
  11. WebArbiter: A Principle-Guided Reasoning Process Reward Model for Web Agents

    Yao Zhang, Shijie Tang, Zeyu Li +2

    cs.AIarXiv:2601.21872v22026
  12. Improving Unsupervised Defect Segmentation by Applying Structural Similarity to Autoencoders

    Paul Bergmann, Sindy Löwe, Michael Fauser +2

    cs.CVcs.LGarXiv:1807.02011v32018
  13. Is Faster R-CNN Doing Well for Pedestrian Detection?

    Liliang Zhang, Liang Lin, Xiaodan Liang +1

    cs.CVarXiv:1607.07032v22016
  14. An Interactive Agent for Requirement-Driven Candidate Sourcing

    Yuanpeng He, Fangjing Li, Xiangyu Ru +10

    cs.SEarXiv:2608.23501v12026
  15. Interp3D: Correspondence-aware Interpolation for Generative Textured 3D Morphing

    Xiaolu Liu, Yicong Li, Qiyuan He +4

    cs.CVarXiv:2601.14103v12026
  16. Single Image Super-Resolution via a Holistic Attention Network

    Ben Niu, Weilei Wen, Wenqi Ren +6

    eess.IVcs.CVarXiv:2008.08767v12020
  17. Using the Output Embedding to Improve Language Models

    Ofir Press, Lior Wolf

    cs.CLarXiv:1608.05859v32016
  18. Large-Small Model Collaboration for Zero-Shot Surgical Phase Recognition

    Yiyi Zhang, Ying Zheng, Wenxin Fan +5

    cs.CVarXiv:2608.22879v12026
  19. NAS-Bench-101: Towards Reproducible Neural Architecture Search

    Chris Ying, Aaron Klein, Esteban Real +3

    cs.LGstat.MLarXiv:1902.09635v22019
  20. Evolving from Tool User to Creator via Training-Free Experience Reuse in Multimodal Reasoning

    Xintian Shen, Jiawei Chen, Lihao Zheng +3

    cs.AIarXiv:2602.01983v12026
  21. Multi-task Sequence to Sequence Learning

    Minh-Thang Luong, Quoc V. Le, Ilya Sutskever +2

    cs.LGcs.CLstat.MLarXiv:1511.06114v42015
  22. Variable Rate Image Compression with Recurrent Neural Networks

    George Toderici, Sean M. O'Malley, Sung Jin Hwang +5

    cs.CVcs.LGcs.NEarXiv:1511.06085v52015
  23. A Threshold Homomorphic Blockchain Architecture for Secure and Scalable IoT Sensor Data Aggregation

    Narendra Kumar Dewangan, Mounira Msahli

    cs.CRcs.NIarXiv:2608.23396v12026
  24. Beyond Low-frequency Information in Graph Convolutional Networks

    Deyu Bo, Xiao Wang, Chuan Shi +1

    cs.LGcs.SIarXiv:2101.00797v12021
  25. The $\mathbf{Y}$-Combinator for LLMs: Solving Long-Context Rot with $λ$-Calculus

    Amartya Roy, Rasul Tutunov, Xiaotong Ji +2

    cs.LGcs.AIarXiv:2603.20105v12026
  26. Spotter: Efficient Urban Visual Localization via Geo-Referenced Facade Landmarks in GPS-Degraded Environments

    Antoni Valls, Jordi Sanchez-Riera

    cs.CVarXiv:2608.23290v12026
  27. WorldVQA: Measuring Atomic World Knowledge in Multimodal Large Language Models

    Runjie Zhou, Youbo Shao, Haoyu Lu +16

    cs.CVcs.LGarXiv:2602.02537v12026
  28. The Composition Theorem for Differential Privacy

    Peter Kairouz, Sewoong Oh, Pramod Viswanath

    cs.DScs.CRcs.ITarXiv:1311.0776v42013
  29. An Attention Enhanced Graph Convolutional LSTM Network for Skeleton-Based Action Recognition

    Chenyang Si, Wentao Chen, Wei Wang +2

    cs.CVarXiv:1902.09130v22019
  30. Latent Chain-of-Thought as Planning: Decoupling Reasoning from Verbalization

    Jiecong Wang, Hao Peng, Chunyang Liu

    cs.AIcs.CLarXiv:2601.21358v22026
  31. UnifiedQA: Crossing Format Boundaries With a Single QA System

    Daniel Khashabi, Sewon Min, Tushar Khot +4

    cs.CLcs.AIarXiv:2005.00700v32020
  32. ScanNet++: A High-Fidelity Dataset of 3D Indoor Scenes

    Chandan Yeshwanth, Yueh-Cheng Liu, Matthias Nießner +1

    cs.CVarXiv:2308.11417v12023
  33. Towards Bridging the Gap between Large-Scale Pretraining and Efficient Finetuning for Humanoid Control

    Weidong Huang, Zhehan Li, Hangxin Liu +3

    cs.ROarXiv:2601.21363v32026
  34. Benchmarking Large Language Models for News Summarization

    Tianyi Zhang, Faisal Ladhak, Esin Durmus +3

    cs.CLcs.AIcs.LGarXiv:2301.13848v12023
  35. Knowledge is Not Enough: Injecting RL Skills for Continual Adaptation

    Pingzhi Tang, Yiding Wang, Muhan Zhang

    cs.LGcs.AIcs.CLarXiv:2601.11258v22026
  36. MovieQA: Understanding Stories in Movies through Question-Answering

    Makarand Tapaswi, Yukun Zhu, Rainer Stiefelhagen +3

    cs.CVcs.CLarXiv:1512.02902v22015
  37. VoxServe: Streaming-Centric Serving System for Speech Language Models

    Keisuke Kamahori, Wei-Tzu Lee, Atindra Jha +4

    cs.LGcs.AIcs.DCarXiv:2602.00269v12026
  38. AttGAN: Facial Attribute Editing by Only Changing What You Want

    Zhenliang He, Wangmeng Zuo, Meina Kan +2

    cs.CVstat.MLarXiv:1711.10678v32017
  39. QuantLRM: Quantization of Large Reasoning Models via Fine-Tuning Signals

    Nan Zhang, Eugene Kwek, Yusen Zhang +4

    cs.LGcs.AIarXiv:2602.02581v12026
  40. Automatic Differentiation Variational Inference

    Alp Kucukelbir, Dustin Tran, Rajesh Ranganath +2

    stat.MLcs.AIcs.LGarXiv:1603.00788v12016
  41. Aligning Agentic World Models via Knowledgeable Experience Learning

    Baochang Ren, Yunzhi Yao, Rui Sun +3

    cs.CLcs.AIcs.CVarXiv:2601.13247v12026
  42. SafeGround: Know When to Trust GUI Grounding Models via Uncertainty Calibration

    Qingni Wang, Yue Fan, Xin Eric Wang

    cs.AIcs.SEarXiv:2602.02419v22026
  43. Spiking Neural Networks for Continuous Control: Neuromorphic Reinforcement Learning in Conventional Computing

    Jessica Hunter, Md Maruf Hossain Shuvo, Krishna Roy

    cs.LGcs.NEarXiv:2608.22729v12026
  44. SEVerA: Verified Synthesis of Self-Evolving Agents

    Debangshu Banerjee, Changming Xu, Eugene Ie +4

    cs.LGcs.PLcs.SEarXiv:2603.25111v22026
  45. Semantic3D.net: A new Large-scale Point Cloud Classification Benchmark

    Timo Hackel, Nikolay Savinov, Lubor Ladicky +3

    cs.CVcs.LGcs.NEarXiv:1704.03847v12017
  46. S2D2: Fast Decoding for Diffusion LLMs via Training-Free Self-Speculation

    Ligong Han, Hao Wang, Han Gao +2

    cs.CLarXiv:2603.25702v22026
  47. A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT

    Yihan Cao, Siyu Li, Yixin Liu +4

    cs.AIcs.CLcs.LGarXiv:2303.04226v12023
  48. Skin Tokens: A Learned Compact Representation for Unified Autoregressive Rigging

    Jia-peng Zhang, Cheng-Feng Pu, Meng-Hao Guo +2

    cs.GRcs.AIarXiv:2602.04805v12026
  49. Beyond Unimodal Shortcuts: MLLMs as Cross-Modal Reasoners for Grounded Named Entity Recognition

    Jinlong Ma, Yu Zhang, Xuefeng Bai +5

    cs.CLarXiv:2602.04486v12026
  50. Engineered 2D Ising interactions on a trapped-ion quantum simulator with hundreds of spins

    Joseph W. Britton, Brian C. Sawyer, Adam C. Keith +5

    quant-phcond-mat.str-elphysics.comp-pharXiv:1204.5789v12012
  51. Salient Object Detection: A Survey

    Ali Borji, Ming-Ming Cheng, Qibin Hou +2

    cs.CVcs.AIq-bio.NCarXiv:1411.5878v62014
  52. DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models

    Hao Liang, Zhengyang Zhao, Meiyi Qiang +22

    cs.LGcs.CLarXiv:2603.26164v12026
  53. DOLFIN: Automated Finite Element Computing

    Anders Logg, Garth N. Wells

    cs.MSmath.NAarXiv:1103.6248v12011
  54. Distributed Prioritized Experience Replay

    Dan Horgan, John Quan, David Budden +4

    cs.LGarXiv:1803.00933v12018
  55. Learning to Commit: Generating Organic Pull Requests via Online Repository Memory

    Mo Li, L. H. Xu, Qitai Tan +2

    cs.SEcs.CLarXiv:2603.26664v12026
  56. The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation

    Chenyu Mu, Xin He, Qu Yang +13

    cs.CVcs.AIarXiv:2601.17737v32026
  57. LoViF 2026 The First Challenge on Unified Removal of Raindrops and Reflections: Methods and Results

    Zewei He, Xi Tong, Yu Chen +49

    cs.CVarXiv:2608.22723v12026
  58. TALLRec: An Effective and Efficient Tuning Framework to Align Large Language Model with Recommendation

    Keqin Bao, Jizhi Zhang, Yang Zhang +3

    cs.IRarXiv:2305.00447v32023
  59. Optimize Surgical Triplet Recognition: A Knowledge-Driven Mixture-of-Experts Solution

    Yiyi Zhang, Yuchen Yuan, Ying Zheng +4

    cs.CVarXiv:2608.22972v12026
  60. SEGCloud: Semantic Segmentation of 3D Point Clouds

    Lyne P. Tchapmi, Christopher B. Choy, Iro Armeni +2

    cs.CVarXiv:1710.07563v12017