Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
20,041 to 20,100 of 61,350
Cyber-Physical Digital Factory Architecture as the Enabler of Disembodied Work
Tero Kaarlela, Ivan Ruchkin, Jose Outeiro +1
cs.HCarXiv:2609.00195v12026MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
Wei Tao, Yucheng Zhou, Yanlin Wang +3
cs.SEcs.AIarXiv:2403.17927v22024A wildland fire model with data assimilation
Jan Mandel, Lynn S. Bennethum, Jonathan D. Beezley +4
math.NAarXiv:0709.0086v22007Improving the Performance of K-Means for Color Quantization
M. Emre Celebi
cs.GRarXiv:1101.0395v12011REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Jian Hu, Jason Klein Liu, Haotian Xu +1
cs.CLcs.LGarXiv:2501.03262v92025Power Grid Vulnerability to Geographically Correlated Failures - Analysis and Control Implications
Andrey Bernstein, Daniel Bienstock, David Hay +2
eess.SYcs.PFmath.OCarXiv:1206.1099v12012A unifying mutual information view of metric learning: cross-entropy vs. pairwise losses
Malik Boudiaf, Jérôme Rony, Imtiaz Masud Ziko +4
cs.LGcs.CVstat.MLarXiv:2003.08983v32020Long-Context Autoregressive Video Modeling with Next-Frame Prediction
Yuchao Gu, Weijia Mao, Mike Zheng Shou
cs.CVarXiv:2503.19325v32025A Survey of Label-noise Representation Learning: Past, Present and Future
Bo Han, Quanming Yao, Tongliang Liu +4
cs.LGarXiv:2011.04406v22020The Privacy-Hallucination Tradeoff in Differentially Private Language Models
Krithika Ramesh, Krishna Pillutla, Danish Pruthi +1
cs.AIcs.CLarXiv:2609.00492v12026A 4D Light-Field Dataset and CNN Architectures for Material Recognition
Ting-Chun Wang, Jun-Yan Zhu, Ebi Hiroaki +3
cs.CVarXiv:1608.06985v12016Stochastic complexity of vectors containing cluster structure
Daniel Nicorici, Olli Yli-Harja, Jaakko Astola
cs.LGcs.ITstat.MLarXiv:2609.00084v12026Fast L1-L2 minimization via a proximal operator
Yifei Lou, Ming Yan
math.OCcs.ITmath.NAarXiv:1609.09530v42016Baichuan-Omni-1.5 Technical Report
Yadong Li, Jun Liu, Tao Zhang +90
cs.CLcs.SDeess.ASarXiv:2501.15368v12025Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning
Haoji Zhang, Xin Gu, Jiawen Li +7
cs.CVarXiv:2508.04416v22025The shifted proper orthogonal decomposition: A mode decomposition for multiple transport phenomena
Julius Reiss, Philipp Schulze, Jörn Sesterhenn +1
math.NAarXiv:1512.01985v32015From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
Mohamed Amine Ferrag, Norbert Tihanyi, Djallel Hamouda +3
cs.CRcs.AIarXiv:2506.23260v22025Scale-free Networks Well Done
Ivan Voitalov, Pim van der Hoorn, Remco van der Hofstad +1
physics.soc-phcs.SIphysics.data-anarXiv:1811.02071v22018Summaries:한국어jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval
Michael Günther, Saba Sturua, Mohammad Kalim Akram +8
cs.AIcs.CLcs.IRarXiv:2506.18902v32025Efficiently inferring community structure in bipartite networks
Daniel B. Larremore, Aaron Clauset, Abigail Z. Jacobs
cs.SIphysics.data-anphysics.soc-pharXiv:1403.2933v22014Kevin: Multi-Turn RL for Generating CUDA Kernels
Carlo Baronio, Pietro Marsella, Ben Pan +2
cs.LGcs.AIcs.PFarXiv:2507.11948v12025A hybrid quantum-classical neural network for learning to route
Marcus Rolf Peter Ritt, Alexsandro Santos da Rosa Júnior, Marcos Vinicius Reballo +2
cs.LGquant-pharXiv:2609.00489v12026A Placement Vulnerability Study in Multi-tenant Public Clouds
Venkatanathan Varadarajan, Yinqian Zhang, Thomas Ristenpart +1
cs.CRarXiv:1507.03114v12015MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining
LLM-Core Xiaomi, :, Bingquan Xia +62
cs.CLcs.AIcs.LGarXiv:2505.07608v22025Large-Scale Multilingual Speech Recognition with a Streaming End-to-End Model
Anjuli Kannan, Arindrima Datta, Tara N. Sainath +6
eess.AScs.LGcs.SDarXiv:1909.05330v12019CAMIE: Co-Engagement-Aware Multimodal Item Embeddings for Snap Dynamic Product Ads Retrieval
Xiaodong Liu, Siman Wang, Congfei Zhang +9
cs.IRarXiv:2608.30255v12026Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
Runpeng Yu, Xinyin Ma, Xinchao Wang
cs.CVarXiv:2505.16990v32025Combining Deep Reinforcement Learning and Search for Imperfect-Information Games
Noam Brown, Anton Bakhtin, Adam Lerer +1
cs.GTcs.AIcs.LGarXiv:2007.13544v22020Back to the Drawing Board: A Critical Evaluation of Poisoning Attacks on Production Federated Learning
Virat Shejwalkar, Amir Houmansadr, Peter Kairouz +1
cs.LGcs.CRcs.DCarXiv:2108.10241v22021Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
Ze Liu, Yutong Lin, Yue Cao +5
cs.CVcs.LGarXiv:2103.14030v22021DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reza Yazdani Aminabadi, Samyam Rajbhandari, Minjia Zhang +8
cs.LGcs.DCcs.PFarXiv:2207.00032v12022Evaluation of Neural Architectures Trained with Square Loss vs Cross-Entropy in Classification Tasks
Like Hui, Mikhail Belkin
cs.LGstat.MLarXiv:2006.07322v52020Fair k-Center Clustering for Data Summarization
Matthäus Kleindessner, Pranjal Awasthi, Jamie Morgenstern
stat.MLcs.DScs.LGarXiv:1901.08628v22019MuJoCo Playground
Kevin Zakka, Baruch Tabanpour, Qiayuan Liao +10
cs.ROarXiv:2502.08844v12025St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World
Haiwen Feng, Junyi Zhang, Qianqian Wang +5
cs.CVarXiv:2504.13152v12025NRTR: A No-Recurrence Sequence-to-Sequence Model For Scene Text Recognition
Fenfen Sheng, Zhineng Chen, Bo Xu
cs.CVarXiv:1806.00926v22018SBERT-WK: A Sentence Embedding Method by Dissecting BERT-based Word Models
Bin Wang, C. -C. Jay Kuo
cs.CLcs.LGcs.MMarXiv:2002.06652v22020WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
Yu Shang, Zhuohang Li, Yiding Ma +18
cs.CVcs.ROarXiv:2602.08971v22026RoboReward: General-Purpose Vision-Language Reward Models for Robotics
Tony Lee, Andrew Wagenmaker, Karl Pertsch +3
cs.ROarXiv:2601.00675v22026Multivariate Industrial Time Series with Cyber-Attack Simulation: Fault Detection Using an LSTM-based Predictive Data Model
Pavel Filonov, Andrey Lavrentyev, Artem Vorontsov
cs.LGstat.MLarXiv:1612.06676v22016PrivBasis: Frequent Itemset Mining with Differential Privacy
Ninghui Li, Wahbeh Qardaji, Dong Su +1
cs.DBarXiv:1208.0093v12012Spec-Driven Development for Agentic Software Engineering: Harnessing Human-Agent Teamwork
Jessica Diaz, Joaquin Gayoso, Andrea Cimminio +1
cs.SEarXiv:2609.00252v12026TransZero: Attribute-guided Transformer for Zero-Shot Learning
Shiming Chen, Ziming Hong, Yang Liu +6
cs.CVcs.AIarXiv:2112.01683v12021PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides
Hao Zheng, Xinyan Guan, Hao Kong +7
cs.AIcs.CLarXiv:2501.03936v32025An Unsupervised Sentence Embedding Method by Mutual Information Maximization
Yan Zhang, Ruidan He, Zuozhu Liu +2
cs.CLcs.LGarXiv:2009.12061v22020LaViDa: A Large Diffusion Language Model for Multimodal Understanding
Shufan Li, Konstantinos Kallidromitis, Hritik Bansal +7
cs.CVarXiv:2505.16839v42025Zero-Shot Respiratory Sound Classification through LLM-Augmented Audio-Text Alignment
Mustafa Talha İlerisoy, Hung Manh Pham, Mathias Funk +2
cs.CLcs.AIcs.SDarXiv:2609.00055v12026Competence-based Multimodal Curriculum Learning for Medical Report Generation
Fenglin Liu, Shen Ge, Yuexian Zou +1
cs.CLcs.CVcs.LGarXiv:2206.14579v32022Adaptive Observer of Nonlinear One-Sided Lipschitz Systems Using Estimated State Regressors With Finite Excitation
Hamid Taghavifar, Brian Delgado Aguilar
eess.SYarXiv:2608.30977v12026Graph Neural News Recommendation with Long-term and Short-term Interest Modeling
Linmei Hu, Chen Li, Chuan Shi +2
cs.IRcs.CLcs.LGarXiv:1910.14025v22019Predicting Successful Memes using Network and Community Structure
Lilian Weng, Filippo Menczer, Yong-Yeol Ahn
cs.SIcs.CYphysics.data-anarXiv:1403.6199v22014Codes in Permutations and Error Correction for Rank Modulation
Alexander Barg, Arya Mazumdar
cs.ITarXiv:0908.4094v32009PentestGPT: An LLM-empowered Automatic Penetration Testing Tool
Gelei Deng, Yi Liu, Víctor Mayoral-Vilches +7
cs.SEcs.CRarXiv:2308.06782v22023Diffusion Adversarial Post-Training for One-Step Video Generation
Shanchuan Lin, Xin Xia, Yuxi Ren +3
cs.CVcs.AIcs.LGarXiv:2501.08316v32025Index-Free Dynamic Edge Retrieval with Energy-Tail-Aware Partial Scans
Mohammad Arif Rasyidi, Omar Alhussein
cs.IRarXiv:2609.01820v12026Summaries:한국어VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models
Xiangdong Zhang, Jiaqi Liao, Shaofeng Zhang +4
cs.CVarXiv:2505.23656v12025From Images to Textual Prompts: Zero-shot VQA with Frozen Large Language Models
Jiaxian Guo, Junnan Li, Dongxu Li +4
cs.CVcs.MMarXiv:2212.10846v32022GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
Qianhui Wu, Kanzhi Cheng, Rui Yang +15
cs.CLcs.AIcs.CVarXiv:2506.03143v12025CAST: Component-Aligned 3D Scene Reconstruction from an RGB Image
Kaixin Yao, Longwen Zhang, Xinhao Yan +6
cs.CVarXiv:2502.12894v22025A note on the reduction from LTLf to LTL
Alexandre Duret-Lutz
cs.FLcs.LOarXiv:2609.00379v12026