Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
52,501 to 52,560 of 61,428
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning
Mingshuang Luo, Shuang Liang, Zhengkun Rong +7
cs.CVcs.AIarXiv:2601.21716v12026An Analysis of Scale Invariance in Object Detection - SNIP
Bharat Singh, Larry S. Davis
cs.CVarXiv:1711.08189v22017The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies
Chenxu Wang, Chaozhuo Li, Songyang Liu +10
cs.CLarXiv:2602.09877v22026Few-Shot Learning via Embedding Adaptation with Set-to-Set Functions
Han-Jia Ye, Hexiang Hu, De-Chuan Zhan +1
cs.LGcs.CVarXiv:1812.03664v62018Controllable blind deblurring with diffusion models
Imane Si Salah, Emile Cribelier, Thomas Veit +2
cs.CVarXiv:2608.23343v12026Syntax Element Encryption for H.265/HEVC Using Chaotic Map-Based Coefficient Scrambling Scheme
Liang-Wei Li, Chung-Nan Lee, Kishu Gupta +2
cs.CRarXiv:2608.22573v12026Neural Operator based Multi-Field Reconstruction of Inner Solar Boundary State
Vignesh Kumar Pandian Sathia, Reza Mansouri, Dustin J. Kempton +2
cs.LGastro-ph.IMastro-ph.SRarXiv:2608.22782v12026When to Memorize and When to Stop: Gated Recurrent Memory for Long-Context Reasoning
Leheng Sheng, Yongtao Zhang, Wenchang Ma +6
cs.CLcs.AIarXiv:2602.10560v12026A new difference scheme for the time fractional diffusion equation
A. A. Alikhanov
math.NAmath-pharXiv:1404.5221v32014CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data
Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4
cs.CLcs.IRcs.LGarXiv:1911.00359v22019MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
MiniCPM Team, Wenhao An, Yingfa Chen +44
cs.CLcs.AIcs.LGarXiv:2602.11761v22026Geometric Autoencoder for Diffusion Models
Hangyu Liu, Jianyong Wang, Yutao Sun
cs.CVarXiv:2603.10365v22026Using Pre-Training Can Improve Model Robustness and Uncertainty
Dan Hendrycks, Kimin Lee, Mantas Mazeika
cs.LGcs.CVstat.MLarXiv:1901.09960v52019Privasis: Synthesizing the Largest "Public" Private Dataset from Scratch
Hyunwoo Kim, Niloofar Mireshghallah, Michael Duan +11
cs.CLcs.AIarXiv:2602.03183v12026Wonder3D: Single Image to 3D using Cross-Domain Diffusion
Xiaoxiao Long, Yuan-Chen Guo, Cheng Lin +8
cs.CVarXiv:2310.15008v32023InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem
Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10
cs.CLcs.AIcs.IRarXiv:2602.14367v22026Physics-informed neural networks with hard constraints for inverse design
Lu Lu, Raphael Pestourie, Wenjie Yao +3
physics.comp-phcs.LGarXiv:2102.04626v12021Attack of the Tails: Yes, You Really Can Backdoor Federated Learning
Hongyi Wang, Kartik Sreenivasan, Shashank Rajput +5
cs.LGcs.CRcs.DCarXiv:2007.05084v12020Cost-Efficient RAG for Entity Matching with LLMs: A Blocking-based Exploration
Chuangtao Ma, Zeyu Zhang, Arijit Khan +2
cs.DBcs.CLarXiv:2602.05708v12026MambaIR: A Simple Baseline for Image Restoration with State-Space Model
Hang Guo, Jinmin Li, Tao Dai +3
cs.CVarXiv:2402.15648v32024Grounding Free-Form Instructions for Fashion Complementary Image Generation
Matteo Attimonelli, Claudio Pomo, Alessandro De Bellis +3
cs.CVarXiv:2608.23302v12026DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts
Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu +2
cs.CVcs.AIcs.LGarXiv:2608.23363v12026Characterizing Necessary Losers to Explain Tournaments Losers
Contet Clément, Umberto Grandi, Jérôme Mengin
cs.AIarXiv:2608.23446v12026The Trinity of Consistency as a Defining Principle for General World Models
Jingxuan Wei, Siyuan Li, Yuhang Xu +21
cs.AIarXiv:2602.23152v12026Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens
Weihao Liu, Dehai Min, Lu Cheng
cs.CLarXiv:2602.10229v22026LoST: Level of Semantics Tokenization for 3D Shapes
Niladri Shekhar Dutt, Zifan Shi, Paul Guerrero +4
cs.CVcs.GRcs.LGarXiv:2603.17995v12026CogVLM: Visual Expert for Pretrained Language Models
Weihan Wang, Qingsong Lv, Wenmeng Yu +13
cs.CVarXiv:2311.03079v22023HopSkipJumpAttack: A Query-Efficient Decision-Based Attack
Jianbo Chen, Michael I. Jordan, Martin J. Wainwright
cs.LGcs.CRmath.OCarXiv:1904.02144v52019SOLO: Segmenting Objects by Locations
Xinlong Wang, Tao Kong, Chunhua Shen +2
cs.CVarXiv:1912.04488v32019Reservoir of Importance: Learning Semi-Structured Sparsity with Differentiable Subset Sampling
Ha Dinh, Xuan Duy Ta, Khoat Than +1
cs.LGarXiv:2608.23048v12026Spanning the Visual Analogy Space with a Weight Basis of LoRAs
Hila Manor, Rinon Gal, Haggai Maron +2
cs.CVcs.AIcs.GRarXiv:2602.15727v22026Molecular LLM Agents: From Architectural Design to Scientific Autonomy
Jiatong Li, Wengyu Zhang, Weida Wang +8
cs.CLcs.AIarXiv:2608.23104v12026Aligning Large Multimodal Models with Factually Augmented RLHF
Zhiqing Sun, Sheng Shen, Shengcao Cao +9
cs.CVcs.CLarXiv:2309.14525v12023VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Zesen Cheng, Sicong Leng, Hang Zhang +8
cs.CVcs.CLarXiv:2406.07476v32024DyaDiT: A Multi-Modal Diffusion Transformer for Socially Favorable Dyadic Gesture Generation
Yichen Peng, Jyun-Ting Song, Siyeol Jung +7
cs.CVarXiv:2602.23165v22026Weakly Supervised Deep Detection Networks
Hakan Bilen, Andrea Vedaldi
cs.CVarXiv:1511.02853v42015Structural Inference in Undocumented Mobile Databases: A Reproducible Benchmark for Evaluating Agentic Reasoning in Digital Forensics
Jeel Piyushkumar Khatiwala, Divyangkumar Patel, Weifeng Xu
cs.CRarXiv:2608.21470v12026CMI-RewardBench: Evaluating Music Reward Models with Compositional Multimodal Instruction
Yinghao Ma, Haiwen Xia, Hewei Gao +9
cs.SDcs.AIcs.LGarXiv:2603.00610v32026Segment Anything Model for Medical Image Analysis: an Experimental Study
Maciej A. Mazurowski, Haoyu Dong, Hanxue Gu +3
cs.CVcs.AIcs.LGarXiv:2304.10517v32023Formalizing and Automating Fine-Grained Move Refactorings Across Methods
Kota Yasuhara, Shinpei Hayashi
cs.SEarXiv:2608.23377v12026Cooperative Non-Orthogonal Multiple Access with Simultaneous Wireless Information and Power Transfer
Yuanwei Liu, Zhiguo Ding, Maged Elkashlan +1
cs.ITarXiv:1511.02833v12015InfoPO: Information-Driven Policy Optimization for User-Centric Agents
Fanqi Kong, Jiayi Zhang, Mingyi Deng +3
cs.AIarXiv:2603.00656v22026A survey of graphical languages for monoidal categories
Peter Selinger
math.CTarXiv:0908.3347v12009LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation
Jiaxiang Tang, Zhaoxi Chen, Xiaokang Chen +3
cs.CVarXiv:2402.05054v12024Contrastive Clustering
Yunfan Li, Peng Hu, Zitao Liu +3
cs.LGcs.CVstat.MLarXiv:2009.09687v12020Future Optical Flow Prediction Improves Robot Control & Video Generation
Kanchana Ranasinghe, Honglu Zhou, Yu Fang +7
cs.CVarXiv:2601.10781v12026The All-Sky Automated Survey for Supernovae (ASAS-SN) Light Curve Server v1.0
C. S. Kochanek, B. J. Shappee, K. Z. Stanek +9
astro-ph.SRastro-ph.IMarXiv:1706.07060v12017Joint Optic Disc and Cup Segmentation Based on Multi-label Deep Network and Polar Transformation
Huazhu Fu, Jun Cheng, Yanwu Xu +3
cs.CVarXiv:1801.00926v32018WnW: Waxing-and-Waning KV Cache for Long-Form Speech LLMs
Yiming Yao, Chenyang Lyu, Xuanfan Ni +4
cs.CLcs.SDarXiv:2608.22704v12026ASVspoof 2019: Future Horizons in Spoofed and Fake Audio Detection
Massimiliano Todisco, Xin Wang, Ville Vestman +7
eess.AScs.CRcs.SDarXiv:1904.05441v22019VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding
Jiapeng Shi, Junke Wang, Zuyao You +2
cs.CVarXiv:2601.07290v12026Spend Search Where It Pays: Value-Guided Structured Sampling and Optimization for Generative Recommendation
Jie Jiang, Yangru Huang, Zeyu Wang +4
cs.AIcs.LGarXiv:2602.10699v22026KILT: a Benchmark for Knowledge Intensive Language Tasks
Fabio Petroni, Aleksandra Piktus, Angela Fan +10
cs.CLcs.AIcs.IRarXiv:2009.02252v42020Cyber-Security in Smart Grid: Survey and Challenges
Zakaria El Mrabet, Hassan El Ghazi, Naima Kaabouch +1
cs.CRcs.NIarXiv:1809.02609v12018Bilateral Multi-Perspective Matching for Natural Language Sentences
Zhiguo Wang, Wael Hamza, Radu Florian
cs.AIcs.CLarXiv:1702.03814v32017Definitional Sensitivity in Media Bias Detection: A Multi-Definition Dataset and Benchmark
Martin Wessel, Timo Spinde, Jürgen Pfeffer +1
cs.CLarXiv:2608.23095v12026Building machines that adapt and compute like brains
Nikolaus Kriegeskorte, Robert M. Mok
cs.AIq-bio.NCarXiv:1711.04203v12017High-Fidelity Audio Compression with Improved RVQGAN
Rithesh Kumar, Prem Seetharaman, Alejandro Luebs +2
cs.SDcs.LGeess.ASarXiv:2306.06546v22023SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generation
Vaibhav Agrawal, Rishubh Parihar, Pradhaan Bhat +2
cs.CVcs.AIarXiv:2602.23359v12026HydroShear: Hydroelastic Shear Simulation for Tactile Sim-to-Real Reinforcement Learning
An Dang, Jayjun Lee, Mustafa Mukadam +4
cs.ROcs.AIarXiv:2603.00446v12026