Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
19,981 to 20,040 of 61,306
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning
Yang Chen, Zhuolin Yang, Zihan Liu +5
cs.LGcs.AIcs.CLarXiv:2505.16400v32025LoGoNet: Towards Accurate 3D Object Detection with Local-to-Global Cross-Modal Fusion
Xin Li, Tao Ma, Yuenan Hou +8
cs.CVarXiv:2303.03595v22023MUSt3R: Multi-view Network for Stereo 3D Reconstruction
Yohann Cabon, Lucas Stoffl, Leonid Antsfeld +4
cs.CVarXiv:2503.01661v12025Probably Approximately Correct MDP Learning and Control With Temporal Logic Constraints
Jie Fu, Ufuk Topcu
eess.SYcs.LGcs.LOarXiv:1404.7073v22014MEt3R: Measuring Multi-View Consistency in Generated Images
Mohammad Asim, Christopher Wewer, Thomas Wimmer +2
cs.CVcs.LGeess.IVarXiv:2501.06336v22025Representation-Free Model Predictive Control for Dynamic Motions in Quadrupeds
Yanran Ding, Abhishek Pandala, Chuanzheng Li +2
cs.ROarXiv:2012.10002v12020Exploit the Connectivity: Multi-Object Tracking with TrackletNet
Gaoang Wang, Yizhou Wang, Haotian Zhang +2
cs.CVarXiv:1811.07258v12018EgoLife: Towards Egocentric Life Assistant
Jingkang Yang, Shuai Liu, Hongming Guo +19
cs.CVarXiv:2503.03803v32025Machine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Evan Crothers, Nathalie Japkowicz, Herna Viktor
cs.CLcs.CRcs.CYarXiv:2210.07321v42022Towards a Reliable and Practical Eval Pipeline
Emma Thuong Nguyen, Abhishek Ghose
cs.AIcs.SEarXiv:2609.00805v12026Various thresholds for $\ell_1$-optimization in compressed sensing
Mihailo Stojnic
cs.ITarXiv:0907.3666v12009Transparency and Explanation in Deep Reinforcement Learning Neural Networks
Rahul Iyer, Yuezhang Li, Huao Li +3
cs.LGstat.MLarXiv:1809.06061v12018Evading Deepfake-Image Detectors with White- and Black-Box Attacks
Nicholas Carlini, Hany Farid
cs.CVcs.CRarXiv:2004.00622v12020Probing Pre-Trained Language Models for Cross-Cultural Differences in Values
Arnav Arora, Lucie-Aimée Kaffee, Isabelle Augenstein
cs.CLarXiv:2203.13722v32022The Open Molecules 2025 (OMol25) Dataset, Evaluations, and Models
Daniel S. Levine, Muhammed Shuaibi, Evan Walter Clark Spotte-Smith +20
physics.chem-pharXiv:2505.08762v22025DreamAvatar: Text-and-Shape Guided 3D Human Avatar Generation via Diffusion Models
Yukang Cao, Yan-Pei Cao, Kai Han +2
cs.CVarXiv:2304.00916v32023Cyber-Physical Digital Factory Architecture as the Enabler of Disembodied Work
Tero Kaarlela, Ivan Ruchkin, Jose Outeiro +1
cs.HCarXiv:2609.00195v12026MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution
Wei Tao, Yucheng Zhou, Yanlin Wang +3
cs.SEcs.AIarXiv:2403.17927v22024A wildland fire model with data assimilation
Jan Mandel, Lynn S. Bennethum, Jonathan D. Beezley +4
math.NAarXiv:0709.0086v22007Improving the Performance of K-Means for Color Quantization
M. Emre Celebi
cs.GRarXiv:1101.0395v12011REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Jian Hu, Jason Klein Liu, Haotian Xu +1
cs.CLcs.LGarXiv:2501.03262v92025Power Grid Vulnerability to Geographically Correlated Failures - Analysis and Control Implications
Andrey Bernstein, Daniel Bienstock, David Hay +2
eess.SYcs.PFmath.OCarXiv:1206.1099v12012A unifying mutual information view of metric learning: cross-entropy vs. pairwise losses
Malik Boudiaf, Jérôme Rony, Imtiaz Masud Ziko +4
cs.LGcs.CVstat.MLarXiv:2003.08983v32020Long-Context Autoregressive Video Modeling with Next-Frame Prediction
Yuchao Gu, Weijia Mao, Mike Zheng Shou
cs.CVarXiv:2503.19325v32025A Survey of Label-noise Representation Learning: Past, Present and Future
Bo Han, Quanming Yao, Tongliang Liu +4
cs.LGarXiv:2011.04406v22020The Privacy-Hallucination Tradeoff in Differentially Private Language Models
Krithika Ramesh, Krishna Pillutla, Danish Pruthi +1
cs.AIcs.CLarXiv:2609.00492v12026A 4D Light-Field Dataset and CNN Architectures for Material Recognition
Ting-Chun Wang, Jun-Yan Zhu, Ebi Hiroaki +3
cs.CVarXiv:1608.06985v12016Stochastic complexity of vectors containing cluster structure
Daniel Nicorici, Olli Yli-Harja, Jaakko Astola
cs.LGcs.ITstat.MLarXiv:2609.00084v12026Fast L1-L2 minimization via a proximal operator
Yifei Lou, Ming Yan
math.OCcs.ITmath.NAarXiv:1609.09530v42016Baichuan-Omni-1.5 Technical Report
Yadong Li, Jun Liu, Tao Zhang +90
cs.CLcs.SDeess.ASarXiv:2501.15368v12025Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning
Haoji Zhang, Xin Gu, Jiawen Li +7
cs.CVarXiv:2508.04416v22025The shifted proper orthogonal decomposition: A mode decomposition for multiple transport phenomena
Julius Reiss, Philipp Schulze, Jörn Sesterhenn +1
math.NAarXiv:1512.01985v32015From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
Mohamed Amine Ferrag, Norbert Tihanyi, Djallel Hamouda +3
cs.CRcs.AIarXiv:2506.23260v22025Scale-free Networks Well Done
Ivan Voitalov, Pim van der Hoorn, Remco van der Hofstad +1
physics.soc-phcs.SIphysics.data-anarXiv:1811.02071v22018Summaries:한국어jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval
Michael Günther, Saba Sturua, Mohammad Kalim Akram +8
cs.AIcs.CLcs.IRarXiv:2506.18902v32025Efficiently inferring community structure in bipartite networks
Daniel B. Larremore, Aaron Clauset, Abigail Z. Jacobs
cs.SIphysics.data-anphysics.soc-pharXiv:1403.2933v22014Kevin: Multi-Turn RL for Generating CUDA Kernels
Carlo Baronio, Pietro Marsella, Ben Pan +2
cs.LGcs.AIcs.PFarXiv:2507.11948v12025A hybrid quantum-classical neural network for learning to route
Marcus Rolf Peter Ritt, Alexsandro Santos da Rosa Júnior, Marcos Vinicius Reballo +2
cs.LGquant-pharXiv:2609.00489v12026A Placement Vulnerability Study in Multi-tenant Public Clouds
Venkatanathan Varadarajan, Yinqian Zhang, Thomas Ristenpart +1
cs.CRarXiv:1507.03114v12015MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining
LLM-Core Xiaomi, :, Bingquan Xia +62
cs.CLcs.AIcs.LGarXiv:2505.07608v22025Large-Scale Multilingual Speech Recognition with a Streaming End-to-End Model
Anjuli Kannan, Arindrima Datta, Tara N. Sainath +6
eess.AScs.LGcs.SDarXiv:1909.05330v12019CAMIE: Co-Engagement-Aware Multimodal Item Embeddings for Snap Dynamic Product Ads Retrieval
Xiaodong Liu, Siman Wang, Congfei Zhang +9
cs.IRarXiv:2608.30255v12026Dimple: Discrete Diffusion Multimodal Large Language Model with Parallel Decoding
Runpeng Yu, Xinyin Ma, Xinchao Wang
cs.CVarXiv:2505.16990v32025Combining Deep Reinforcement Learning and Search for Imperfect-Information Games
Noam Brown, Anton Bakhtin, Adam Lerer +1
cs.GTcs.AIcs.LGarXiv:2007.13544v22020Back to the Drawing Board: A Critical Evaluation of Poisoning Attacks on Production Federated Learning
Virat Shejwalkar, Amir Houmansadr, Peter Kairouz +1
cs.LGcs.CRcs.DCarXiv:2108.10241v22021Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
Ze Liu, Yutong Lin, Yue Cao +5
cs.CVcs.LGarXiv:2103.14030v22021DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale
Reza Yazdani Aminabadi, Samyam Rajbhandari, Minjia Zhang +8
cs.LGcs.DCcs.PFarXiv:2207.00032v12022Evaluation of Neural Architectures Trained with Square Loss vs Cross-Entropy in Classification Tasks
Like Hui, Mikhail Belkin
cs.LGstat.MLarXiv:2006.07322v52020Fair k-Center Clustering for Data Summarization
Matthäus Kleindessner, Pranjal Awasthi, Jamie Morgenstern
stat.MLcs.DScs.LGarXiv:1901.08628v22019MuJoCo Playground
Kevin Zakka, Baruch Tabanpour, Qiayuan Liao +10
cs.ROarXiv:2502.08844v12025St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World
Haiwen Feng, Junyi Zhang, Qianqian Wang +5
cs.CVarXiv:2504.13152v12025NRTR: A No-Recurrence Sequence-to-Sequence Model For Scene Text Recognition
Fenfen Sheng, Zhineng Chen, Bo Xu
cs.CVarXiv:1806.00926v22018SBERT-WK: A Sentence Embedding Method by Dissecting BERT-based Word Models
Bin Wang, C. -C. Jay Kuo
cs.CLcs.LGcs.MMarXiv:2002.06652v22020WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
Yu Shang, Zhuohang Li, Yiding Ma +18
cs.CVcs.ROarXiv:2602.08971v22026RoboReward: General-Purpose Vision-Language Reward Models for Robotics
Tony Lee, Andrew Wagenmaker, Karl Pertsch +3
cs.ROarXiv:2601.00675v22026Multivariate Industrial Time Series with Cyber-Attack Simulation: Fault Detection Using an LSTM-based Predictive Data Model
Pavel Filonov, Andrey Lavrentyev, Artem Vorontsov
cs.LGstat.MLarXiv:1612.06676v22016PrivBasis: Frequent Itemset Mining with Differential Privacy
Ninghui Li, Wahbeh Qardaji, Dong Su +1
cs.DBarXiv:1208.0093v12012Spec-Driven Development for Agentic Software Engineering: Harnessing Human-Agent Teamwork
Jessica Diaz, Joaquin Gayoso, Andrea Cimminio +1
cs.SEarXiv:2609.00252v12026TransZero: Attribute-guided Transformer for Zero-Shot Learning
Shiming Chen, Ziming Hong, Yang Liu +6
cs.CVcs.AIarXiv:2112.01683v12021PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides
Hao Zheng, Xinyan Guan, Hao Kong +7
cs.AIcs.CLarXiv:2501.03936v32025