Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
51,241 to 51,300 of 61,350
VideoDetective: Clue Hunting via both Extrinsic Query and Intrinsic Relevance for Long Video Understanding
Ruoliu Yang, Chu Wu, Caifeng Shan +2
cs.CVarXiv:2603.22285v22026WorldCache: Content-Aware Caching for Accelerated Video World Models
Umair Nawaz, Ahmed Heakl, Ufaq Khan +3
cs.CVcs.AIcs.CLarXiv:2603.22286v12026Asymmetric Contextual Modulation for Infrared Small Target Detection
Yimian Dai, Yiquan Wu, Fei Zhou +1
cs.CVarXiv:2009.14530v12020SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
Yuan Liu, Cheng Lin, Zijiao Zeng +4
cs.CVcs.AIcs.GRarXiv:2309.03453v22023CayleyNets: Graph Convolutional Neural Networks with Complex Rational Spectral Filters
Ron Levie, Federico Monti, Xavier Bresson +1
cs.LGarXiv:1705.07664v22017SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning
Haoyu Huang, Jinfa Huang, Zhongwei Wan +3
cs.CVcs.CLarXiv:2603.23483v22026Net2Net: Accelerating Learning via Knowledge Transfer
Tianqi Chen, Ian Goodfellow, Jonathon Shlens
cs.LGarXiv:1511.05641v42015Robust Reasoning Benchmark
Pavel Golikov, Evgenii Opryshko, Gennady Pekhimenko +1
cs.LGcs.AIcs.CLarXiv:2604.08571v32026Transformer Meets Tracker: Exploiting Temporal Context for Robust Visual Tracking
Ning Wang, Wengang Zhou, Jie Wang +1
cs.CVarXiv:2103.11681v22021Mining Educational Data to Analyze Students' Performance
Brijesh Kumar Baradwaj, Saurabh Pal
cs.IRarXiv:1201.3417v12012STRIDE: When to Speak Meets Sequence Denoising for Streaming Video Understanding
Junho Kim, Hosu Lee, James M. Rehg +2
cs.CVcs.AIarXiv:2603.27593v12026Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation
Xue Liu, Xin Ma, Yuxin Ma +36
cs.AIcs.CLarXiv:2604.02368v42026Going Deeper With Directly-Trained Larger Spiking Neural Networks
Hanle Zheng, Yujie Wu, Lei Deng +2
cs.NEcs.AIarXiv:2011.05280v22020Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems
Shanglin Wu, Yuyang Luo, Yueqing Liang +4
cs.MAcs.AIarXiv:2604.03295v12026ScaleNet: An Unsupervised Representation Learning Method for Limited Information
Huili Huang, M. Mahdi Roozbahani
cs.CVarXiv:2310.02386v12023LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training
Andreas Hochlehnert, Marianna Nezhurina, Mehdi Cherti +9
cs.CVcs.AIcs.LGarXiv:2608.24845v12026Mish: A Self Regularized Non-Monotonic Activation Function
Diganta Misra
cs.LGcs.CVcs.NEarXiv:1908.08681v32019Scalable Training of Artificial Neural Networks with Adaptive Sparse Connectivity inspired by Network Science
Decebal Constantin Mocanu, Elena Mocanu, Peter Stone +3
cs.NEcs.AIcs.LGarXiv:1707.04780v22017Semantic Segmentation using Adversarial Networks
Pauline Luc, Camille Couprie, Soumith Chintala +1
cs.CVarXiv:1611.08408v12016Genie: Generative Interactive Environments
Jake Bruce, Michael Dennis, Ashley Edwards +22
cs.LGcs.AIcs.CVarXiv:2402.15391v12024AnimalCLAP: Taxonomy-Aware Language-Audio Pretraining for Species Recognition and Trait Inference
Risa Shinoda, Kaede Shiohara, Nakamasa Inoue +2
cs.SDcs.LGarXiv:2603.22053v12026Advanced LLM-Enhanced Intent-Based 5G Network Management using Dynamic Semantic Routes
Thomas Benton Townsend, Dimitrios Michael Manias
cs.NIcs.LGeess.SYarXiv:2608.22644v12026VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Changhan Wang, Morgane Rivière, Ann Lee +6
cs.CLeess.ASarXiv:2101.00390v22021Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
WonJun Moon, Hyun Seok Seong, Jae-Pil Heo
cs.CVcs.LGarXiv:2603.22758v12026Frustratingly Simple Few-Shot Object Detection
Xin Wang, Thomas E. Huang, Trevor Darrell +2
cs.CVarXiv:2003.06957v12020Know3D: Prompting 3D Generation with Knowledge from Vision-Language Models
Wenyue Chen, Wenjue Chen, Peng Li +6
cs.CVarXiv:2603.22782v22026Neural Approaches to Conversational AI
Jianfeng Gao, Michel Galley, Lihong Li
cs.CLarXiv:1809.08267v32018SIMART: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLM
Chuanrui Zhang, Minghan Qin, Yuang Wang +3
cs.CVcs.GRcs.ROarXiv:2603.23386v12026UniFunc3D: Unified Active Spatial-Temporal Grounding for 3D Functionality Segmentation
Jiaying Lin, Dan Xu
cs.CVarXiv:2603.23478v12026ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation
Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein
cs.AIcs.CLcs.IRarXiv:2608.22559v12026PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders
Niccolò Cavagnero, Narges Norouzi, Gijs Dubbelman +1
cs.CVarXiv:2603.25398v12026Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos
Abdullah Hamdi, Changchun Yang, Xin Gao
eess.IVcs.CVcs.HCarXiv:2603.25645v32026ThreatLens: Evidence-Guided Ranking of High-Priority CVEs
Soroush Motamedi Sedeh, Panteha Shahrivar, Malaika Qureshi +2
cs.CRarXiv:2608.22306v12026Tabular foundation models for non-tabular tasks
Goran Nakerst, John Brennan, Wouter Beugeling +1
cs.LGarXiv:2608.22594v12026ViGoR-Bench: How Far Are Visual Generative Models From Zero-Shot Visual Reasoners?
Haonan Han, Jiancheng Huang, Xiaopeng Sun +7
cs.CVcs.AIarXiv:2603.25823v12026Unsupervised Label Noise Modeling and Loss Correction
Eric Arazo, Diego Ortego, Paul Albert +2
cs.CVarXiv:1904.11238v22019Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping
Jesse Dodge, Gabriel Ilharco, Roy Schwartz +3
cs.CLcs.LGarXiv:2002.06305v12020Millimeter Wave Vehicular Communication to Support Massive Automotive Sensing
Junil Choi, Vutha Va, Nuria Gonzalez-Prelcic +3
cs.ITarXiv:1602.06456v22016The Design and Implementation of XiaoIce, an Empathetic Social Chatbot
Li Zhou, Jianfeng Gao, Di Li +1
cs.HCcs.AIcs.CLarXiv:1812.08989v22018PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning
Shaoxuan Li, Zhixuan Zhao, Hanze Deng +9
cs.CVcs.AIcs.CLarXiv:2603.26653v12026TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets
Zhixuan Liu, Peter Schaldenbrand, Yijun Li +5
cs.CVarXiv:2603.27520v12026Emergent Social Intelligence Risks in Generative Multi-Agent Systems
Yue Huang, Yu Jiang, Wenjie Wang +12
cs.MAcs.CLcs.CYarXiv:2603.27771v22026Ghost-FWL: A Large-Scale Full-Waveform LiDAR Dataset for Ghost Detection and Removal
Kazuma Ikeda, Ryosei Hara, Rokuto Nagata +6
cs.CVarXiv:2603.28224v12026WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report
Junjie Zhou, Ke Mei, Lei Li +3
cs.CVcs.CLcs.IRarXiv:2608.24053v12026Summaries:한국어Tensor networks for complex quantum systems
Roman Orus
cond-mat.str-elhep-latquant-pharXiv:1812.04011v22018DScribe: Library of Descriptors for Machine Learning in Materials Science
Lauri Himanen, Marc O. J. Jäger, Eiaki V. Morooka +5
cond-mat.mtrl-scics.LGarXiv:1904.08875v12019Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models
Itay Yona, Dan Barzilay, Michael Karasik +1
cs.CLcs.AIarXiv:2604.01404v22026Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects
Jonathan Tremblay, Thang To, Balakumar Sundaralingam +3
cs.ROarXiv:1809.10790v12018Spider-Sense: Intrinsic Risk Sensing for Efficient Agent Defense with Hierarchical Adaptive Screening
Zhenxiong Yu, Zhi Yang, Zhiheng Jin +19
cs.CRcs.AIarXiv:2602.05386v22026MemRerank: Preference Memory for Personalized Product Reranking
Zhiyuan Peng, Xuyang Wu, Huaixiao Tou +2
cs.CLcs.AIcs.LGarXiv:2603.29247v32026Think Anywhere in Code Generation
Xue Jiang, Tianyu Zhang, Ge Li +8
cs.SEcs.LGarXiv:2603.29957v32026Implicit Neural Representation Facilitates Unified Universal Vision Encoding
Matthew Gwilliam, Xiao Wang, Xuefeng Hu +1
cs.CVarXiv:2601.14256v12026MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal Large Language Models
Sangyun Chung, Se Yeon Kim, Youngchae Chee +1
cs.AIarXiv:2601.21181v12026Ebisu: Benchmarking Large Language Models in Japanese Finance
Xueqing Peng, Ruoyu Xiang, Fan Zhang +9
cs.CLarXiv:2602.01479v12026Multi-digit Number Recognition from Street View Imagery using Deep Convolutional Neural Networks
Ian J. Goodfellow, Yaroslav Bulatov, Julian Ibarz +2
cs.CVarXiv:1312.6082v42013Summaries:한국어WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
Haipeng Luo, Qingfeng Sun, Can Xu +8
cs.CLcs.AIcs.LGarXiv:2308.09583v32023Hybrid Panels: Toward Human-AI Collaboration in Survey Research
Julia Romberg, Tobias Gummer, Gabriella Lapesa +2
cs.CLcs.AIcs.CYarXiv:2608.22582v12026Visual Memory Injection Attacks for Multi-Turn Conversations
Christian Schlarmann, Matthias Hein
cs.CVcs.LGarXiv:2602.15927v12026Influence and Passivity in Social Media
Daniel M. Romero, Wojciech Galuba, Sitaram Asur +1
cs.CYphysics.soc-pharXiv:1008.1253v12010PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval
Tianyi Xu, Rong Shan, Junjie Wu +11
cs.IRcs.AIcs.CVarXiv:2603.01493v22026