Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

51,241 to 51,300 of 61,350

  1. VideoDetective: Clue Hunting via both Extrinsic Query and Intrinsic Relevance for Long Video Understanding

    Ruoliu Yang, Chu Wu, Caifeng Shan +2

    cs.CVarXiv:2603.22285v22026
  2. WorldCache: Content-Aware Caching for Accelerated Video World Models

    Umair Nawaz, Ahmed Heakl, Ufaq Khan +3

    cs.CVcs.AIcs.CLarXiv:2603.22286v12026
  3. Asymmetric Contextual Modulation for Infrared Small Target Detection

    Yimian Dai, Yiquan Wu, Fei Zhou +1

    cs.CVarXiv:2009.14530v12020
  4. SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

    Yuan Liu, Cheng Lin, Zijiao Zeng +4

    cs.CVcs.AIcs.GRarXiv:2309.03453v22023
  5. CayleyNets: Graph Convolutional Neural Networks with Complex Rational Spectral Filters

    Ron Levie, Federico Monti, Xavier Bresson +1

    cs.LGarXiv:1705.07664v22017
  6. SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

    Haoyu Huang, Jinfa Huang, Zhongwei Wan +3

    cs.CVcs.CLarXiv:2603.23483v22026
  7. Net2Net: Accelerating Learning via Knowledge Transfer

    Tianqi Chen, Ian Goodfellow, Jonathon Shlens

    cs.LGarXiv:1511.05641v42015
  8. Robust Reasoning Benchmark

    Pavel Golikov, Evgenii Opryshko, Gennady Pekhimenko +1

    cs.LGcs.AIcs.CLarXiv:2604.08571v32026
  9. Transformer Meets Tracker: Exploiting Temporal Context for Robust Visual Tracking

    Ning Wang, Wengang Zhou, Jie Wang +1

    cs.CVarXiv:2103.11681v22021
  10. Mining Educational Data to Analyze Students' Performance

    Brijesh Kumar Baradwaj, Saurabh Pal

    cs.IRarXiv:1201.3417v12012
  11. STRIDE: When to Speak Meets Sequence Denoising for Streaming Video Understanding

    Junho Kim, Hosu Lee, James M. Rehg +2

    cs.CVcs.AIarXiv:2603.27593v12026
  12. Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation

    Xue Liu, Xin Ma, Yuxin Ma +36

    cs.AIcs.CLarXiv:2604.02368v42026
  13. Going Deeper With Directly-Trained Larger Spiking Neural Networks

    Hanle Zheng, Yujie Wu, Lei Deng +2

    cs.NEcs.AIarXiv:2011.05280v22020
  14. Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems

    Shanglin Wu, Yuyang Luo, Yueqing Liang +4

    cs.MAcs.AIarXiv:2604.03295v12026
  15. ScaleNet: An Unsupervised Representation Learning Method for Limited Information

    Huili Huang, M. Mahdi Roozbahani

    cs.CVarXiv:2310.02386v12023
  16. LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training

    Andreas Hochlehnert, Marianna Nezhurina, Mehdi Cherti +9

    cs.CVcs.AIcs.LGarXiv:2608.24845v12026
  17. Mish: A Self Regularized Non-Monotonic Activation Function

    Diganta Misra

    cs.LGcs.CVcs.NEarXiv:1908.08681v32019
  18. Scalable Training of Artificial Neural Networks with Adaptive Sparse Connectivity inspired by Network Science

    Decebal Constantin Mocanu, Elena Mocanu, Peter Stone +3

    cs.NEcs.AIcs.LGarXiv:1707.04780v22017
  19. Semantic Segmentation using Adversarial Networks

    Pauline Luc, Camille Couprie, Soumith Chintala +1

    cs.CVarXiv:1611.08408v12016
  20. Genie: Generative Interactive Environments

    Jake Bruce, Michael Dennis, Ashley Edwards +22

    cs.LGcs.AIcs.CVarXiv:2402.15391v12024
  21. AnimalCLAP: Taxonomy-Aware Language-Audio Pretraining for Species Recognition and Trait Inference

    Risa Shinoda, Kaede Shiohara, Nakamasa Inoue +2

    cs.SDcs.LGarXiv:2603.22053v12026
  22. Advanced LLM-Enhanced Intent-Based 5G Network Management using Dynamic Semantic Routes

    Thomas Benton Townsend, Dimitrios Michael Manias

    cs.NIcs.LGeess.SYarXiv:2608.22644v12026
  23. VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

    Changhan Wang, Morgane Rivière, Ann Lee +6

    cs.CLeess.ASarXiv:2101.00390v22021
  24. Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning

    WonJun Moon, Hyun Seok Seong, Jae-Pil Heo

    cs.CVcs.LGarXiv:2603.22758v12026
  25. Frustratingly Simple Few-Shot Object Detection

    Xin Wang, Thomas E. Huang, Trevor Darrell +2

    cs.CVarXiv:2003.06957v12020
  26. Know3D: Prompting 3D Generation with Knowledge from Vision-Language Models

    Wenyue Chen, Wenjue Chen, Peng Li +6

    cs.CVarXiv:2603.22782v22026
  27. Neural Approaches to Conversational AI

    Jianfeng Gao, Michel Galley, Lihong Li

    cs.CLarXiv:1809.08267v32018
  28. SIMART: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLM

    Chuanrui Zhang, Minghan Qin, Yuang Wang +3

    cs.CVcs.GRcs.ROarXiv:2603.23386v12026
  29. UniFunc3D: Unified Active Spatial-Temporal Grounding for 3D Functionality Segmentation

    Jiaying Lin, Dan Xu

    cs.CVarXiv:2603.23478v12026
  30. ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation

    Kaustubh D. Dhole, Charles L. A. Clarke, Eugene Y. Agichtein

    cs.AIcs.CLcs.IRarXiv:2608.22559v12026
  31. PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders

    Niccolò Cavagnero, Narges Norouzi, Gijs Dubbelman +1

    cs.CVarXiv:2603.25398v12026
  32. Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos

    Abdullah Hamdi, Changchun Yang, Xin Gao

    eess.IVcs.CVcs.HCarXiv:2603.25645v32026
  33. ThreatLens: Evidence-Guided Ranking of High-Priority CVEs

    Soroush Motamedi Sedeh, Panteha Shahrivar, Malaika Qureshi +2

    cs.CRarXiv:2608.22306v12026
  34. Tabular foundation models for non-tabular tasks

    Goran Nakerst, John Brennan, Wouter Beugeling +1

    cs.LGarXiv:2608.22594v12026
  35. ViGoR-Bench: How Far Are Visual Generative Models From Zero-Shot Visual Reasoners?

    Haonan Han, Jiancheng Huang, Xiaopeng Sun +7

    cs.CVcs.AIarXiv:2603.25823v12026
  36. Unsupervised Label Noise Modeling and Loss Correction

    Eric Arazo, Diego Ortego, Paul Albert +2

    cs.CVarXiv:1904.11238v22019
  37. Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

    Jesse Dodge, Gabriel Ilharco, Roy Schwartz +3

    cs.CLcs.LGarXiv:2002.06305v12020
  38. Millimeter Wave Vehicular Communication to Support Massive Automotive Sensing

    Junil Choi, Vutha Va, Nuria Gonzalez-Prelcic +3

    cs.ITarXiv:1602.06456v22016
  39. The Design and Implementation of XiaoIce, an Empathetic Social Chatbot

    Li Zhou, Jianfeng Gao, Di Li +1

    cs.HCcs.AIcs.CLarXiv:1812.08989v22018
  40. PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning

    Shaoxuan Li, Zhixuan Zhao, Hanze Deng +9

    cs.CVcs.AIcs.CLarXiv:2603.26653v12026
  41. TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets

    Zhixuan Liu, Peter Schaldenbrand, Yijun Li +5

    cs.CVarXiv:2603.27520v12026
  42. Emergent Social Intelligence Risks in Generative Multi-Agent Systems

    Yue Huang, Yu Jiang, Wenjie Wang +12

    cs.MAcs.CLcs.CYarXiv:2603.27771v22026
  43. Ghost-FWL: A Large-Scale Full-Waveform LiDAR Dataset for Ghost Detection and Removal

    Kazuma Ikeda, Ryosei Hara, Rokuto Nagata +6

    cs.CVarXiv:2603.28224v12026
  44. WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report

    Junjie Zhou, Ke Mei, Lei Li +3

    cs.CVcs.CLcs.IRarXiv:2608.24053v12026
    Summaries:한국어
  45. Tensor networks for complex quantum systems

    Roman Orus

    cond-mat.str-elhep-latquant-pharXiv:1812.04011v22018
  46. DScribe: Library of Descriptors for Machine Learning in Materials Science

    Lauri Himanen, Marc O. J. Jäger, Eiaki V. Morooka +5

    cond-mat.mtrl-scics.LGarXiv:1904.08875v12019
  47. Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models

    Itay Yona, Dan Barzilay, Michael Karasik +1

    cs.CLcs.AIarXiv:2604.01404v22026
  48. Deep Object Pose Estimation for Semantic Robotic Grasping of Household Objects

    Jonathan Tremblay, Thang To, Balakumar Sundaralingam +3

    cs.ROarXiv:1809.10790v12018
  49. Spider-Sense: Intrinsic Risk Sensing for Efficient Agent Defense with Hierarchical Adaptive Screening

    Zhenxiong Yu, Zhi Yang, Zhiheng Jin +19

    cs.CRcs.AIarXiv:2602.05386v22026
  50. MemRerank: Preference Memory for Personalized Product Reranking

    Zhiyuan Peng, Xuyang Wu, Huaixiao Tou +2

    cs.CLcs.AIcs.LGarXiv:2603.29247v32026
  51. Think Anywhere in Code Generation

    Xue Jiang, Tianyu Zhang, Ge Li +8

    cs.SEcs.LGarXiv:2603.29957v32026
  52. Implicit Neural Representation Facilitates Unified Universal Vision Encoding

    Matthew Gwilliam, Xiao Wang, Xuefeng Hu +1

    cs.CVarXiv:2601.14256v12026
  53. MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal Large Language Models

    Sangyun Chung, Se Yeon Kim, Youngchae Chee +1

    cs.AIarXiv:2601.21181v12026
  54. Ebisu: Benchmarking Large Language Models in Japanese Finance

    Xueqing Peng, Ruoyu Xiang, Fan Zhang +9

    cs.CLarXiv:2602.01479v12026
  55. Multi-digit Number Recognition from Street View Imagery using Deep Convolutional Neural Networks

    Ian J. Goodfellow, Yaroslav Bulatov, Julian Ibarz +2

    cs.CVarXiv:1312.6082v42013
    Summaries:한국어
  56. WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

    Haipeng Luo, Qingfeng Sun, Can Xu +8

    cs.CLcs.AIcs.LGarXiv:2308.09583v32023
  57. Hybrid Panels: Toward Human-AI Collaboration in Survey Research

    Julia Romberg, Tobias Gummer, Gabriella Lapesa +2

    cs.CLcs.AIcs.CYarXiv:2608.22582v12026
  58. Visual Memory Injection Attacks for Multi-Turn Conversations

    Christian Schlarmann, Matthias Hein

    cs.CVcs.LGarXiv:2602.15927v12026
  59. Influence and Passivity in Social Media

    Daniel M. Romero, Wojciech Galuba, Sitaram Asur +1

    cs.CYphysics.soc-pharXiv:1008.1253v12010
  60. PhotoBench: Beyond Visual Matching Towards Personalized Intent-Driven Photo Retrieval

    Tianyi Xu, Rong Shan, Junjie Wu +11

    cs.IRcs.AIcs.CVarXiv:2603.01493v22026