Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
52,441 to 52,500 of 61,393
CogAgent: A Visual Language Model for GUI Agents
Wenyi Hong, Weihan Wang, Qingsong Lv +11
cs.CVarXiv:2312.08914v32023Data-Enabled Predictive Control: In the Shallows of the DeePC
Jeremy Coulson, John Lygeros, Florian Dörfler
math.OCarXiv:1811.05890v22018GSAR: Goal-State-Anchor Rewards for Mobile GUI Agents with Self-Evolving Data Synthesis
Long Zhang, Yuhan Chen, Chaoran Zhang +7
cs.AIarXiv:2608.22847v12026High resolution dynamical mapping of social interactions with active RFID
Alain Barrat, Ciro Cattuto, Vittoria Colizza +3
cs.CYcs.HCphysics.soc-pharXiv:0811.4170v22008MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering
Chuanzhe Guo, Jingjing Wu, Sijun He +10
cs.SEcs.AIarXiv:2601.22859v32026Linguistic Knowledge and Transferability of Contextual Representations
Nelson F. Liu, Matt Gardner, Yonatan Belinkov +2
cs.CLarXiv:1903.08855v52019MediSkill-Evo: Process-Constrained Self-Evolution for Evidence-Grounded Clinical Interaction
Ruoyu Wu, Shenfu Xie, Yinqian Sun +2
cs.AIarXiv:2608.23397v12026DICE: Diffusion Large Language Models Excel at Generating CUDA Kernels
Haolei Bai, Lingcheng Kong, Xueyi Chen +3
cs.LGcs.CLarXiv:2602.11715v22026HateXplain: A Benchmark Dataset for Explainable Hate Speech Detection
Binny Mathew, Punyajoy Saha, Seid Muhie Yimam +3
cs.CLcs.AIcs.SIarXiv:2012.10289v22020AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution Traces
Sungho Park, Wonjoong Kim, Rongyuan Tan +10
cs.AIcs.CLcs.LGarXiv:2608.23041v12026Efficient Off-Policy Meta-Reinforcement Learning via Probabilistic Context Variables
Kate Rakelly, Aurick Zhou, Deirdre Quillen +2
cs.LGcs.AIstat.MLarXiv:1903.08254v12019PointFlow: 3D Point Cloud Generation with Continuous Normalizing Flows
Guandao Yang, Xun Huang, Zekun Hao +3
cs.CVcs.LGarXiv:1906.12320v32019A Physical Response-and-Memory Model for Muon Optimization
Yinze Hu, Hongjun Xiang, Xingao Gong +1
cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2608.22994v12026MOReL : Model-Based Offline Reinforcement Learning
Rahul Kidambi, Aravind Rajeswaran, Praneeth Netrapalli +1
cs.LGcs.AIstat.MLarXiv:2005.05951v32020LaViDa-R1: Advancing Reasoning for Unified Multimodal Diffusion Language Models
Shufan Li, Yuchen Zhu, Jiuxiang Gu +6
cs.CVarXiv:2602.14147v22026Revisiting Spatial-Temporal Similarity: A Deep Learning Framework for Traffic Prediction
Huaxiu Yao, Xianfeng Tang, Hua Wei +2
cs.LGarXiv:1803.01254v22018Understanding vs. Generation: Navigating Optimization Dilemma in Multimodal Models
Sen Ye, Mengde Xu, Shuyang Gu +3
cs.CVcs.AIarXiv:2602.15772v22026A Rewriting System for Convex Optimization Problems
Akshay Agrawal, Robin Verschueren, Steven Diamond +1
math.OCcs.MSarXiv:1709.04494v22017Knowing Isn't Understanding: Re-grounding Generative Proactivity with Epistemic and Behavioral Insight
Kirandeep Kaur, Xingda Lyu, Chirag Shah
cs.CYcs.AIcs.LGarXiv:2602.15259v22026Learning Modality-Specific Representations with Self-Supervised Multi-Task Learning for Multimodal Sentiment Analysis
Wenmeng Yu, Hua Xu, Ziqi Yuan +1
cs.CLarXiv:2102.04830v12021Adaptive Restart for Accelerated Gradient Schemes
Brendan O'Donoghue, Emmanuel Candes
math.OCarXiv:1204.3982v12012TRACE: A Self-Evolving Skill Bank for Consistent, Limit-Aware LLM Agents
Wenhao Wu, Menghao Zhang, Xin Wang +3
cs.CLcs.AIarXiv:2608.22793v12026Modalities Should Talk to Each Other: Dual-Stream Multimodal Learning for Long-Horizon Influenza Forecasting
Seyed Mohammad Hossein Hashemi, Mohsen Hooshmand, Parvin Razzaghi
cs.AIstat.AParXiv:2608.23373v12026Multi-Scale Progressive Fusion Network for Single Image Deraining
Kui Jiang, Zhongyuan Wang, Peng Yi +5
cs.CVcs.LGeess.IVarXiv:2003.10985v22020Quantum Error Correction for Beginners
Simon J. Devitt, Kae Nemoto, William J. Munro
quant-pharXiv:0905.2794v42009DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning
Mingshuang Luo, Shuang Liang, Zhengkun Rong +7
cs.CVcs.AIarXiv:2601.21716v12026An Analysis of Scale Invariance in Object Detection - SNIP
Bharat Singh, Larry S. Davis
cs.CVarXiv:1711.08189v22017The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies
Chenxu Wang, Chaozhuo Li, Songyang Liu +10
cs.CLarXiv:2602.09877v22026Few-Shot Learning via Embedding Adaptation with Set-to-Set Functions
Han-Jia Ye, Hexiang Hu, De-Chuan Zhan +1
cs.LGcs.CVarXiv:1812.03664v62018Controllable blind deblurring with diffusion models
Imane Si Salah, Emile Cribelier, Thomas Veit +2
cs.CVarXiv:2608.23343v12026Syntax Element Encryption for H.265/HEVC Using Chaotic Map-Based Coefficient Scrambling Scheme
Liang-Wei Li, Chung-Nan Lee, Kishu Gupta +2
cs.CRarXiv:2608.22573v12026Neural Operator based Multi-Field Reconstruction of Inner Solar Boundary State
Vignesh Kumar Pandian Sathia, Reza Mansouri, Dustin J. Kempton +2
cs.LGastro-ph.IMastro-ph.SRarXiv:2608.22782v12026When to Memorize and When to Stop: Gated Recurrent Memory for Long-Context Reasoning
Leheng Sheng, Yongtao Zhang, Wenchang Ma +6
cs.CLcs.AIarXiv:2602.10560v12026A new difference scheme for the time fractional diffusion equation
A. A. Alikhanov
math.NAmath-pharXiv:1404.5221v32014CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data
Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4
cs.CLcs.IRcs.LGarXiv:1911.00359v22019MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
MiniCPM Team, Wenhao An, Yingfa Chen +44
cs.CLcs.AIcs.LGarXiv:2602.11761v22026Geometric Autoencoder for Diffusion Models
Hangyu Liu, Jianyong Wang, Yutao Sun
cs.CVarXiv:2603.10365v22026Using Pre-Training Can Improve Model Robustness and Uncertainty
Dan Hendrycks, Kimin Lee, Mantas Mazeika
cs.LGcs.CVstat.MLarXiv:1901.09960v52019Privasis: Synthesizing the Largest "Public" Private Dataset from Scratch
Hyunwoo Kim, Niloofar Mireshghallah, Michael Duan +11
cs.CLcs.AIarXiv:2602.03183v12026Wonder3D: Single Image to 3D using Cross-Domain Diffusion
Xiaoxiao Long, Yuan-Chen Guo, Cheng Lin +8
cs.CVarXiv:2310.15008v32023InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem
Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10
cs.CLcs.AIcs.IRarXiv:2602.14367v22026Physics-informed neural networks with hard constraints for inverse design
Lu Lu, Raphael Pestourie, Wenjie Yao +3
physics.comp-phcs.LGarXiv:2102.04626v12021Attack of the Tails: Yes, You Really Can Backdoor Federated Learning
Hongyi Wang, Kartik Sreenivasan, Shashank Rajput +5
cs.LGcs.CRcs.DCarXiv:2007.05084v12020Cost-Efficient RAG for Entity Matching with LLMs: A Blocking-based Exploration
Chuangtao Ma, Zeyu Zhang, Arijit Khan +2
cs.DBcs.CLarXiv:2602.05708v12026MambaIR: A Simple Baseline for Image Restoration with State-Space Model
Hang Guo, Jinmin Li, Tao Dai +3
cs.CVarXiv:2402.15648v32024Grounding Free-Form Instructions for Fashion Complementary Image Generation
Matteo Attimonelli, Claudio Pomo, Alessandro De Bellis +3
cs.CVarXiv:2608.23302v12026DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts
Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu +2
cs.CVcs.AIcs.LGarXiv:2608.23363v12026Characterizing Necessary Losers to Explain Tournaments Losers
Contet Clément, Umberto Grandi, Jérôme Mengin
cs.AIarXiv:2608.23446v12026The Trinity of Consistency as a Defining Principle for General World Models
Jingxuan Wei, Siyuan Li, Yuhang Xu +21
cs.AIarXiv:2602.23152v12026Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens
Weihao Liu, Dehai Min, Lu Cheng
cs.CLarXiv:2602.10229v22026LoST: Level of Semantics Tokenization for 3D Shapes
Niladri Shekhar Dutt, Zifan Shi, Paul Guerrero +4
cs.CVcs.GRcs.LGarXiv:2603.17995v12026CogVLM: Visual Expert for Pretrained Language Models
Weihan Wang, Qingsong Lv, Wenmeng Yu +13
cs.CVarXiv:2311.03079v22023HopSkipJumpAttack: A Query-Efficient Decision-Based Attack
Jianbo Chen, Michael I. Jordan, Martin J. Wainwright
cs.LGcs.CRmath.OCarXiv:1904.02144v52019SOLO: Segmenting Objects by Locations
Xinlong Wang, Tao Kong, Chunhua Shen +2
cs.CVarXiv:1912.04488v32019Reservoir of Importance: Learning Semi-Structured Sparsity with Differentiable Subset Sampling
Ha Dinh, Xuan Duy Ta, Khoat Than +1
cs.LGarXiv:2608.23048v12026Spanning the Visual Analogy Space with a Weight Basis of LoRAs
Hila Manor, Rinon Gal, Haggai Maron +2
cs.CVcs.AIcs.GRarXiv:2602.15727v22026Molecular LLM Agents: From Architectural Design to Scientific Autonomy
Jiatong Li, Wengyu Zhang, Weida Wang +8
cs.CLcs.AIarXiv:2608.23104v12026Aligning Large Multimodal Models with Factually Augmented RLHF
Zhiqing Sun, Sheng Shen, Shengcao Cao +9
cs.CVcs.CLarXiv:2309.14525v12023VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Zesen Cheng, Sicong Leng, Hang Zhang +8
cs.CVcs.CLarXiv:2406.07476v32024DyaDiT: A Multi-Modal Diffusion Transformer for Socially Favorable Dyadic Gesture Generation
Yichen Peng, Jyun-Ting Song, Siyeol Jung +7
cs.CVarXiv:2602.23165v22026