Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
20,461 to 20,520 of 61,350
Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities
Ziwei Zhou, Rui Wang, Zuxuan Wu +1
cs.AIcs.CLcs.CVarXiv:2505.17862v22025Communication Lower Bounds for Statistical Estimation Problems via a Distributed Data Processing Inequality
Mark Braverman, Ankit Garg, Tengyu Ma +2
cs.LGcs.CCcs.ITarXiv:1506.07216v32015Domain Control for Neural Machine Translation
Catherine Kobus, Josep Crego, Jean Senellart
cs.CLarXiv:1612.06140v22016Astraea: A Decentralized Blockchain Oracle
John Adler, Ryan Berryhill, Andreas Veneris +3
cs.CRarXiv:1808.00528v12018Focus on Local: Detecting Lane Marker from Bottom Up via Key Point
Zhan Qu, Huan Jin, Yang Zhou +2
cs.CVarXiv:2105.13680v12021An Emerging NVM-Based On-Chip Training Architecture with Non-Ideality Mitigation Through Bipolar Weight Distributions
Peng Dang, Youna Huang, Yintao He +1
cs.ARarXiv:2609.01948v12026RecipeQA: A Challenge Dataset for Multimodal Comprehension of Cooking Recipes
Semih Yagcioglu, Aykut Erdem, Erkut Erdem +1
cs.CLcs.CVarXiv:1809.00812v12018Synergistic Information Disentanglement for Omni-modal Slide Representation Learning in Computational Pathology
Mingxin Liu, Chengfei Cai, Anwen Lu +5
cs.CVarXiv:2609.02118v12026Detoxifying Toxic Communication: A Design Science Approach to Responsible AI
Hossein Arshadi Soufiani, Henry M. Kim, Hjalmar Turesson +2
cs.CYcs.CLarXiv:2609.00361v12026On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub
Miku Watanabe, Hao Li, Yutaro Kashiwa +3
cs.SEarXiv:2509.14745v32025MemOS: A Memory OS for AI System
Zhiyu Li, Chenyang Xi, Chunyu Li +36
cs.CLarXiv:2507.03724v42025Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
Yunhong Lu, Yanhong Zeng, Haobo Li +9
cs.CVarXiv:2512.04678v22025Beyond Attention or Similarity: Maximizing Conditional Diversity for Token Pruning in MLLMs
Qizhe Zhang, Mengzhen Liu, Lichen Li +5
cs.CVcs.AIarXiv:2506.10967v22025How Important is Weight Symmetry in Backpropagation?
Qianli Liao, Joel Z. Leibo, Tomaso Poggio
cs.LGarXiv:1510.05067v42015Personalized Federated Learning with Feature Alignment and Classifier Collaboration
Jian Xu, Xinyi Tong, Shao-Lun Huang
cs.LGcs.DCarXiv:2306.11867v12023Pinching-Antenna Systems (PASS): A Tutorial
Yuanwei Liu, Hao Jiang, Xiaoxia Xu +8
eess.SParXiv:2508.07572v42025Candidate Generation and Definition-Guided Verification for Sentence-Level Depression Symptom Recognition
Weiming Li, Catarina Barata, Miguel Constante +1
cs.CLarXiv:2609.01833v12026Classifying, Segmenting, and Tracking Object Instances in Video with Mask Propagation
Gedas Bertasius, Lorenzo Torresani
cs.CVarXiv:1912.04573v42019MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling
MiroMind Team, Song Bai, Lidong Bing +52
cs.CLarXiv:2511.11793v32025OTFS Channel Estimation Utilizing Sparse Bayesian Generative Modelling
Louis Anseaume, Benedikt Böck, Franz Weißer +1
eess.SParXiv:2609.01074v12026Faster Retrieval with a Two-Pass Dynamic-Time-Warping Lower Bound
Daniel Lemire
cs.DBcs.CVarXiv:0811.3301v22008Deep Learning Approach to Diabetic Retinopathy Detection
Borys Tymchenko, Philip Marchenko, Dmitry Spodarets
cs.LGstat.MLarXiv:2003.02261v12020Step1X-3D: Towards High-Fidelity and Controllable Generation of Textured 3D Assets
Weiyu Li, Xuanyang Zhang, Zheng Sun +15
cs.CVarXiv:2505.07747v12025Organizing, Orchestrating, and Benchmarking Agent Skills at Ecosystem Scale
Hao Li, Chunjiang Mu, Jianhao Chen +5
cs.CLarXiv:2603.02176v12026Video Super-resolution with Temporal Group Attention
Takashi Isobe, Songjiang Li, Xu Jia +6
cs.CVarXiv:2007.10595v12020OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation
Qijun Gan, Ruizi Yang, Jianke Zhu +2
cs.CVcs.AIcs.MMarXiv:2506.18866v12025Automated Paper Screening for Clinical Reviews Using Large Language Models
Eddie Guo, Mehul Gupta, Jiawen Deng +3
cs.CLcs.AIarXiv:2305.00844v12023TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems
Shaina Raza, Ranjan Sapkota, Manoj Karkee +1
cs.AIarXiv:2506.04133v52025Jointly Reinforcing Diversity and Quality in Language Model Generations
Tianjian Li, Yiming Zhang, Ping Yu +5
cs.CLcs.LGarXiv:2509.02534v12025House of Graphs: a database of interesting graphs
Gunnar Brinkmann, Kris Coolsaet, Jan Goedgebeur +1
math.COcs.DMarXiv:1204.3549v22012Potential-Guided Particle Steering for Negation-Constrained Dexterous Grasping
Geonho Kim, SooGon Kim, Jongmin Lee
cs.ROcs.CVarXiv:2609.00555v12026Learning to Navigate the Energy Landscape
Julien Valentin, Angela Dai, Matthias Nießner +4
cs.CVarXiv:1603.05772v12016LLM Agents for Education: Advances and Applications
Zhendong Chu, Shen Wang, Jian Xie +8
cs.CYcs.AIcs.CLarXiv:2503.11733v22025SipMask: Spatial Information Preservation for Fast Image and Video Instance Segmentation
Jiale Cao, Rao Muhammad Anwer, Hisham Cholakkal +3
cs.CVarXiv:2007.14772v12020VisualPRM: An Effective Process Reward Model for Multimodal Reasoning
Weiyun Wang, Zhangwei Gao, Lianjie Chen +12
cs.CVcs.CLarXiv:2503.10291v12025MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers
Zhenting Wang, Qi Chang, Hemani Patel +8
cs.CLarXiv:2508.20453v12025Curriculum Adversarial Training
Qi-Zhi Cai, Min Du, Chang Liu +1
cs.LGcs.CRstat.MLarXiv:1805.04807v12018Multi-task Mid-level Feature Alignment Network for Unsupervised Cross-Dataset Person Re-Identification
Shan Lin, Haoliang Li, Chang-Tsun Li +1
cs.CVarXiv:1807.01440v22018Training-Time Action Conditioning for Efficient Real-Time Chunking
Kevin Black, Allen Z. Ren, Michael Equi +1
cs.ROcs.AIarXiv:2512.05964v22025Transfer in Deep Reinforcement Learning Using Successor Features and Generalised Policy Improvement
André Barreto, Diana Borsa, John Quan +6
cs.LGcs.AIarXiv:1901.10964v12019Probabilistic Logic Neural Networks for Reasoning
Meng Qu, Jian Tang
cs.LGcs.AIstat.MLarXiv:1906.08495v22019Looped Transformers under the Jacobian Lens: Does the Global Workspace Survive Recurrence?
Wenlong Wang, Fergal Reid
cs.AIarXiv:2609.01924v12026Coreference Resolution as Query-based Span Prediction
Wei Wu, Fei Wang, Arianna Yuan +2
cs.CLarXiv:1911.01746v42019PCGRL: Procedural Content Generation via Reinforcement Learning
Ahmed Khalifa, Philip Bontrager, Sam Earle +1
cs.LGcs.AIstat.MLarXiv:2001.09212v32020ViewSpatial-Bench: Evaluating Multi-perspective Spatial Localization in Vision-Language Models
Dingming Li, Hongxing Li, Zixuan Wang +9
cs.CVcs.AIcs.CLarXiv:2505.21500v22025Behind the Curtain: Learning Occluded Shapes for 3D Object Detection
Qiangeng Xu, Yiqi Zhong, Ulrich Neumann
cs.CVcs.AIcs.LGarXiv:2112.02205v12021AdaWorld: Learning Adaptable World Models with Latent Actions
Shenyuan Gao, Siyuan Zhou, Yilun Du +2
cs.AIcs.CVcs.LGarXiv:2503.18938v42025All You Need is Beyond a Good Init: Exploring Better Solution for Training Extremely Deep Convolutional Neural Networks with Orthonormality and Modulation
Di Xie, Jiang Xiong, Shiliang Pu
cs.CVcs.LGcs.NEarXiv:1703.01827v32017Learning Multi-dimensional Edge Feature-based AU Relation Graph for Facial Action Unit Recognition
Cheng Luo, Siyang Song, Weicheng Xie +2
cs.CVcs.AIarXiv:2205.01782v22022No Task Left Behind: Isotropic Model Merging with Common and Task-Specific Subspaces
Daniel Marczak, Simone Magistri, Sebastian Cygert +3
cs.LGarXiv:2502.04959v32025Learning Mixed Graphical Models
Jason D. Lee, Trevor J. Hastie
stat.MLcs.CVcs.LGarXiv:1205.5012v32012MEM: Multi-Scale Embodied Memory for Vision Language Action Models
Marcel Torne, Karl Pertsch, Homer Walke +14
cs.ROcs.LGarXiv:2603.03596v22026Optimal Uniform Pricing for Multi-Interval Dispatch without Make-Whole Uplifts
Valentina Norambuena-Guzman, Cong Chen, Lang Tong +1
eess.SYecon.EMarXiv:2609.00541v12026Plan-Structured Deep Neural Network Models for Query Performance Prediction
Ryan Marcus, Olga Papaemmanouil
cs.DBarXiv:1902.00132v12019Don't You Know, Pump it Up! Investigating Cryptocurrency Manipulation in Telegram-Driven Activity
Filipe Moura, Giordano Paoletti, Carlos H. G Ferreira +1
cs.SIcs.CEcs.CYarXiv:2609.01176v12026DeeperLab: Single-Shot Image Parser
Tien-Ju Yang, Maxwell D. Collins, Yukun Zhu +6
cs.CVarXiv:1902.05093v22019SkillWeaver: Web Agents can Self-Improve by Discovering and Honing Skills
Boyuan Zheng, Michael Y. Fatemi, Xiaolong Jin +8
cs.AIcs.CLcs.CVarXiv:2504.07079v12025Regularizing Deep Networks with Semantic Data Augmentation
Yulin Wang, Gao Huang, Shiji Song +3
cs.CVcs.LGarXiv:2007.10538v52020Cross-Modal Guidance for Out-of-View Object Search in Simulated Prosthetic Vision
Adyah Rastogi, Apurv Varshney, Tobias Höllerer +1
cs.HCarXiv:2609.01438v12026Sequence Set Design With Good Correlation Properties via Majorization-Minimization
Junxiao Song, Prabhu Babu, Daniel P. Palomar
math.OCarXiv:1510.01899v12015