Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

20,461 to 20,520 of 61,350

  1. Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities

    Ziwei Zhou, Rui Wang, Zuxuan Wu +1

    cs.AIcs.CLcs.CVarXiv:2505.17862v22025
  2. Communication Lower Bounds for Statistical Estimation Problems via a Distributed Data Processing Inequality

    Mark Braverman, Ankit Garg, Tengyu Ma +2

    cs.LGcs.CCcs.ITarXiv:1506.07216v32015
  3. Domain Control for Neural Machine Translation

    Catherine Kobus, Josep Crego, Jean Senellart

    cs.CLarXiv:1612.06140v22016
  4. Astraea: A Decentralized Blockchain Oracle

    John Adler, Ryan Berryhill, Andreas Veneris +3

    cs.CRarXiv:1808.00528v12018
  5. Focus on Local: Detecting Lane Marker from Bottom Up via Key Point

    Zhan Qu, Huan Jin, Yang Zhou +2

    cs.CVarXiv:2105.13680v12021
  6. An Emerging NVM-Based On-Chip Training Architecture with Non-Ideality Mitigation Through Bipolar Weight Distributions

    Peng Dang, Youna Huang, Yintao He +1

    cs.ARarXiv:2609.01948v12026
  7. RecipeQA: A Challenge Dataset for Multimodal Comprehension of Cooking Recipes

    Semih Yagcioglu, Aykut Erdem, Erkut Erdem +1

    cs.CLcs.CVarXiv:1809.00812v12018
  8. Synergistic Information Disentanglement for Omni-modal Slide Representation Learning in Computational Pathology

    Mingxin Liu, Chengfei Cai, Anwen Lu +5

    cs.CVarXiv:2609.02118v12026
  9. Detoxifying Toxic Communication: A Design Science Approach to Responsible AI

    Hossein Arshadi Soufiani, Henry M. Kim, Hjalmar Turesson +2

    cs.CYcs.CLarXiv:2609.00361v12026
  10. On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub

    Miku Watanabe, Hao Li, Yutaro Kashiwa +3

    cs.SEarXiv:2509.14745v32025
  11. MemOS: A Memory OS for AI System

    Zhiyu Li, Chenyang Xi, Chunyu Li +36

    cs.CLarXiv:2507.03724v42025
  12. Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation

    Yunhong Lu, Yanhong Zeng, Haobo Li +9

    cs.CVarXiv:2512.04678v22025
  13. Beyond Attention or Similarity: Maximizing Conditional Diversity for Token Pruning in MLLMs

    Qizhe Zhang, Mengzhen Liu, Lichen Li +5

    cs.CVcs.AIarXiv:2506.10967v22025
  14. How Important is Weight Symmetry in Backpropagation?

    Qianli Liao, Joel Z. Leibo, Tomaso Poggio

    cs.LGarXiv:1510.05067v42015
  15. Personalized Federated Learning with Feature Alignment and Classifier Collaboration

    Jian Xu, Xinyi Tong, Shao-Lun Huang

    cs.LGcs.DCarXiv:2306.11867v12023
  16. Pinching-Antenna Systems (PASS): A Tutorial

    Yuanwei Liu, Hao Jiang, Xiaoxia Xu +8

    eess.SParXiv:2508.07572v42025
  17. Candidate Generation and Definition-Guided Verification for Sentence-Level Depression Symptom Recognition

    Weiming Li, Catarina Barata, Miguel Constante +1

    cs.CLarXiv:2609.01833v12026
  18. Classifying, Segmenting, and Tracking Object Instances in Video with Mask Propagation

    Gedas Bertasius, Lorenzo Torresani

    cs.CVarXiv:1912.04573v42019
  19. MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

    MiroMind Team, Song Bai, Lidong Bing +52

    cs.CLarXiv:2511.11793v32025
  20. OTFS Channel Estimation Utilizing Sparse Bayesian Generative Modelling

    Louis Anseaume, Benedikt Böck, Franz Weißer +1

    eess.SParXiv:2609.01074v12026
  21. Faster Retrieval with a Two-Pass Dynamic-Time-Warping Lower Bound

    Daniel Lemire

    cs.DBcs.CVarXiv:0811.3301v22008
  22. Deep Learning Approach to Diabetic Retinopathy Detection

    Borys Tymchenko, Philip Marchenko, Dmitry Spodarets

    cs.LGstat.MLarXiv:2003.02261v12020
  23. Step1X-3D: Towards High-Fidelity and Controllable Generation of Textured 3D Assets

    Weiyu Li, Xuanyang Zhang, Zheng Sun +15

    cs.CVarXiv:2505.07747v12025
  24. Organizing, Orchestrating, and Benchmarking Agent Skills at Ecosystem Scale

    Hao Li, Chunjiang Mu, Jianhao Chen +5

    cs.CLarXiv:2603.02176v12026
  25. Video Super-resolution with Temporal Group Attention

    Takashi Isobe, Songjiang Li, Xu Jia +6

    cs.CVarXiv:2007.10595v12020
  26. OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation

    Qijun Gan, Ruizi Yang, Jianke Zhu +2

    cs.CVcs.AIcs.MMarXiv:2506.18866v12025
  27. Automated Paper Screening for Clinical Reviews Using Large Language Models

    Eddie Guo, Mehul Gupta, Jiawen Deng +3

    cs.CLcs.AIarXiv:2305.00844v12023
  28. TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems

    Shaina Raza, Ranjan Sapkota, Manoj Karkee +1

    cs.AIarXiv:2506.04133v52025
  29. Jointly Reinforcing Diversity and Quality in Language Model Generations

    Tianjian Li, Yiming Zhang, Ping Yu +5

    cs.CLcs.LGarXiv:2509.02534v12025
  30. House of Graphs: a database of interesting graphs

    Gunnar Brinkmann, Kris Coolsaet, Jan Goedgebeur +1

    math.COcs.DMarXiv:1204.3549v22012
  31. Potential-Guided Particle Steering for Negation-Constrained Dexterous Grasping

    Geonho Kim, SooGon Kim, Jongmin Lee

    cs.ROcs.CVarXiv:2609.00555v12026
  32. Learning to Navigate the Energy Landscape

    Julien Valentin, Angela Dai, Matthias Nießner +4

    cs.CVarXiv:1603.05772v12016
  33. LLM Agents for Education: Advances and Applications

    Zhendong Chu, Shen Wang, Jian Xie +8

    cs.CYcs.AIcs.CLarXiv:2503.11733v22025
  34. SipMask: Spatial Information Preservation for Fast Image and Video Instance Segmentation

    Jiale Cao, Rao Muhammad Anwer, Hisham Cholakkal +3

    cs.CVarXiv:2007.14772v12020
  35. VisualPRM: An Effective Process Reward Model for Multimodal Reasoning

    Weiyun Wang, Zhangwei Gao, Lianjie Chen +12

    cs.CVcs.CLarXiv:2503.10291v12025
  36. MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers

    Zhenting Wang, Qi Chang, Hemani Patel +8

    cs.CLarXiv:2508.20453v12025
  37. Curriculum Adversarial Training

    Qi-Zhi Cai, Min Du, Chang Liu +1

    cs.LGcs.CRstat.MLarXiv:1805.04807v12018
  38. Multi-task Mid-level Feature Alignment Network for Unsupervised Cross-Dataset Person Re-Identification

    Shan Lin, Haoliang Li, Chang-Tsun Li +1

    cs.CVarXiv:1807.01440v22018
  39. Training-Time Action Conditioning for Efficient Real-Time Chunking

    Kevin Black, Allen Z. Ren, Michael Equi +1

    cs.ROcs.AIarXiv:2512.05964v22025
  40. Transfer in Deep Reinforcement Learning Using Successor Features and Generalised Policy Improvement

    André Barreto, Diana Borsa, John Quan +6

    cs.LGcs.AIarXiv:1901.10964v12019
  41. Probabilistic Logic Neural Networks for Reasoning

    Meng Qu, Jian Tang

    cs.LGcs.AIstat.MLarXiv:1906.08495v22019
  42. Looped Transformers under the Jacobian Lens: Does the Global Workspace Survive Recurrence?

    Wenlong Wang, Fergal Reid

    cs.AIarXiv:2609.01924v12026
  43. Coreference Resolution as Query-based Span Prediction

    Wei Wu, Fei Wang, Arianna Yuan +2

    cs.CLarXiv:1911.01746v42019
  44. PCGRL: Procedural Content Generation via Reinforcement Learning

    Ahmed Khalifa, Philip Bontrager, Sam Earle +1

    cs.LGcs.AIstat.MLarXiv:2001.09212v32020
  45. ViewSpatial-Bench: Evaluating Multi-perspective Spatial Localization in Vision-Language Models

    Dingming Li, Hongxing Li, Zixuan Wang +9

    cs.CVcs.AIcs.CLarXiv:2505.21500v22025
  46. Behind the Curtain: Learning Occluded Shapes for 3D Object Detection

    Qiangeng Xu, Yiqi Zhong, Ulrich Neumann

    cs.CVcs.AIcs.LGarXiv:2112.02205v12021
  47. AdaWorld: Learning Adaptable World Models with Latent Actions

    Shenyuan Gao, Siyuan Zhou, Yilun Du +2

    cs.AIcs.CVcs.LGarXiv:2503.18938v42025
  48. All You Need is Beyond a Good Init: Exploring Better Solution for Training Extremely Deep Convolutional Neural Networks with Orthonormality and Modulation

    Di Xie, Jiang Xiong, Shiliang Pu

    cs.CVcs.LGcs.NEarXiv:1703.01827v32017
  49. Learning Multi-dimensional Edge Feature-based AU Relation Graph for Facial Action Unit Recognition

    Cheng Luo, Siyang Song, Weicheng Xie +2

    cs.CVcs.AIarXiv:2205.01782v22022
  50. No Task Left Behind: Isotropic Model Merging with Common and Task-Specific Subspaces

    Daniel Marczak, Simone Magistri, Sebastian Cygert +3

    cs.LGarXiv:2502.04959v32025
  51. Learning Mixed Graphical Models

    Jason D. Lee, Trevor J. Hastie

    stat.MLcs.CVcs.LGarXiv:1205.5012v32012
  52. MEM: Multi-Scale Embodied Memory for Vision Language Action Models

    Marcel Torne, Karl Pertsch, Homer Walke +14

    cs.ROcs.LGarXiv:2603.03596v22026
  53. Optimal Uniform Pricing for Multi-Interval Dispatch without Make-Whole Uplifts

    Valentina Norambuena-Guzman, Cong Chen, Lang Tong +1

    eess.SYecon.EMarXiv:2609.00541v12026
  54. Plan-Structured Deep Neural Network Models for Query Performance Prediction

    Ryan Marcus, Olga Papaemmanouil

    cs.DBarXiv:1902.00132v12019
  55. Don't You Know, Pump it Up! Investigating Cryptocurrency Manipulation in Telegram-Driven Activity

    Filipe Moura, Giordano Paoletti, Carlos H. G Ferreira +1

    cs.SIcs.CEcs.CYarXiv:2609.01176v12026
  56. DeeperLab: Single-Shot Image Parser

    Tien-Ju Yang, Maxwell D. Collins, Yukun Zhu +6

    cs.CVarXiv:1902.05093v22019
  57. SkillWeaver: Web Agents can Self-Improve by Discovering and Honing Skills

    Boyuan Zheng, Michael Y. Fatemi, Xiaolong Jin +8

    cs.AIcs.CLcs.CVarXiv:2504.07079v12025
  58. Regularizing Deep Networks with Semantic Data Augmentation

    Yulin Wang, Gao Huang, Shiji Song +3

    cs.CVcs.LGarXiv:2007.10538v52020
  59. Cross-Modal Guidance for Out-of-View Object Search in Simulated Prosthetic Vision

    Adyah Rastogi, Apurv Varshney, Tobias Höllerer +1

    cs.HCarXiv:2609.01438v12026
  60. Sequence Set Design With Good Correlation Properties via Majorization-Minimization

    Junxiao Song, Prabhu Babu, Daniel P. Palomar

    math.OCarXiv:1510.01899v12015