Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

52,501 to 52,560 of 61,428

  1. DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning

    Mingshuang Luo, Shuang Liang, Zhengkun Rong +7

    cs.CVcs.AIarXiv:2601.21716v12026
  2. An Analysis of Scale Invariance in Object Detection - SNIP

    Bharat Singh, Larry S. Davis

    cs.CVarXiv:1711.08189v22017
  3. The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies

    Chenxu Wang, Chaozhuo Li, Songyang Liu +10

    cs.CLarXiv:2602.09877v22026
  4. Few-Shot Learning via Embedding Adaptation with Set-to-Set Functions

    Han-Jia Ye, Hexiang Hu, De-Chuan Zhan +1

    cs.LGcs.CVarXiv:1812.03664v62018
  5. Controllable blind deblurring with diffusion models

    Imane Si Salah, Emile Cribelier, Thomas Veit +2

    cs.CVarXiv:2608.23343v12026
  6. Syntax Element Encryption for H.265/HEVC Using Chaotic Map-Based Coefficient Scrambling Scheme

    Liang-Wei Li, Chung-Nan Lee, Kishu Gupta +2

    cs.CRarXiv:2608.22573v12026
  7. Neural Operator based Multi-Field Reconstruction of Inner Solar Boundary State

    Vignesh Kumar Pandian Sathia, Reza Mansouri, Dustin J. Kempton +2

    cs.LGastro-ph.IMastro-ph.SRarXiv:2608.22782v12026
  8. When to Memorize and When to Stop: Gated Recurrent Memory for Long-Context Reasoning

    Leheng Sheng, Yongtao Zhang, Wenchang Ma +6

    cs.CLcs.AIarXiv:2602.10560v12026
  9. A new difference scheme for the time fractional diffusion equation

    A. A. Alikhanov

    math.NAmath-pharXiv:1404.5221v32014
  10. CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

    Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4

    cs.CLcs.IRcs.LGarXiv:1911.00359v22019
  11. MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling

    MiniCPM Team, Wenhao An, Yingfa Chen +44

    cs.CLcs.AIcs.LGarXiv:2602.11761v22026
  12. Geometric Autoencoder for Diffusion Models

    Hangyu Liu, Jianyong Wang, Yutao Sun

    cs.CVarXiv:2603.10365v22026
  13. Using Pre-Training Can Improve Model Robustness and Uncertainty

    Dan Hendrycks, Kimin Lee, Mantas Mazeika

    cs.LGcs.CVstat.MLarXiv:1901.09960v52019
  14. Privasis: Synthesizing the Largest "Public" Private Dataset from Scratch

    Hyunwoo Kim, Niloofar Mireshghallah, Michael Duan +11

    cs.CLcs.AIarXiv:2602.03183v12026
  15. Wonder3D: Single Image to 3D using Cross-Domain Diffusion

    Xiaoxiao Long, Yuan-Chen Guo, Cheng Lin +8

    cs.CVarXiv:2310.15008v32023
  16. InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem

    Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10

    cs.CLcs.AIcs.IRarXiv:2602.14367v22026
  17. Physics-informed neural networks with hard constraints for inverse design

    Lu Lu, Raphael Pestourie, Wenjie Yao +3

    physics.comp-phcs.LGarXiv:2102.04626v12021
  18. Attack of the Tails: Yes, You Really Can Backdoor Federated Learning

    Hongyi Wang, Kartik Sreenivasan, Shashank Rajput +5

    cs.LGcs.CRcs.DCarXiv:2007.05084v12020
  19. Cost-Efficient RAG for Entity Matching with LLMs: A Blocking-based Exploration

    Chuangtao Ma, Zeyu Zhang, Arijit Khan +2

    cs.DBcs.CLarXiv:2602.05708v12026
  20. MambaIR: A Simple Baseline for Image Restoration with State-Space Model

    Hang Guo, Jinmin Li, Tao Dai +3

    cs.CVarXiv:2402.15648v32024
  21. Grounding Free-Form Instructions for Fashion Complementary Image Generation

    Matteo Attimonelli, Claudio Pomo, Alessandro De Bellis +3

    cs.CVarXiv:2608.23302v12026
  22. DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts

    Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu +2

    cs.CVcs.AIcs.LGarXiv:2608.23363v12026
  23. Characterizing Necessary Losers to Explain Tournaments Losers

    Contet Clément, Umberto Grandi, Jérôme Mengin

    cs.AIarXiv:2608.23446v12026
  24. The Trinity of Consistency as a Defining Principle for General World Models

    Jingxuan Wei, Siyuan Li, Yuhang Xu +21

    cs.AIarXiv:2602.23152v12026
  25. Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens

    Weihao Liu, Dehai Min, Lu Cheng

    cs.CLarXiv:2602.10229v22026
  26. LoST: Level of Semantics Tokenization for 3D Shapes

    Niladri Shekhar Dutt, Zifan Shi, Paul Guerrero +4

    cs.CVcs.GRcs.LGarXiv:2603.17995v12026
  27. CogVLM: Visual Expert for Pretrained Language Models

    Weihan Wang, Qingsong Lv, Wenmeng Yu +13

    cs.CVarXiv:2311.03079v22023
  28. HopSkipJumpAttack: A Query-Efficient Decision-Based Attack

    Jianbo Chen, Michael I. Jordan, Martin J. Wainwright

    cs.LGcs.CRmath.OCarXiv:1904.02144v52019
  29. SOLO: Segmenting Objects by Locations

    Xinlong Wang, Tao Kong, Chunhua Shen +2

    cs.CVarXiv:1912.04488v32019
  30. Reservoir of Importance: Learning Semi-Structured Sparsity with Differentiable Subset Sampling

    Ha Dinh, Xuan Duy Ta, Khoat Than +1

    cs.LGarXiv:2608.23048v12026
  31. Spanning the Visual Analogy Space with a Weight Basis of LoRAs

    Hila Manor, Rinon Gal, Haggai Maron +2

    cs.CVcs.AIcs.GRarXiv:2602.15727v22026
  32. Molecular LLM Agents: From Architectural Design to Scientific Autonomy

    Jiatong Li, Wengyu Zhang, Weida Wang +8

    cs.CLcs.AIarXiv:2608.23104v12026
  33. Aligning Large Multimodal Models with Factually Augmented RLHF

    Zhiqing Sun, Sheng Shen, Shengcao Cao +9

    cs.CVcs.CLarXiv:2309.14525v12023
  34. VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

    Zesen Cheng, Sicong Leng, Hang Zhang +8

    cs.CVcs.CLarXiv:2406.07476v32024
  35. DyaDiT: A Multi-Modal Diffusion Transformer for Socially Favorable Dyadic Gesture Generation

    Yichen Peng, Jyun-Ting Song, Siyeol Jung +7

    cs.CVarXiv:2602.23165v22026
  36. Weakly Supervised Deep Detection Networks

    Hakan Bilen, Andrea Vedaldi

    cs.CVarXiv:1511.02853v42015
  37. Structural Inference in Undocumented Mobile Databases: A Reproducible Benchmark for Evaluating Agentic Reasoning in Digital Forensics

    Jeel Piyushkumar Khatiwala, Divyangkumar Patel, Weifeng Xu

    cs.CRarXiv:2608.21470v12026
  38. CMI-RewardBench: Evaluating Music Reward Models with Compositional Multimodal Instruction

    Yinghao Ma, Haiwen Xia, Hewei Gao +9

    cs.SDcs.AIcs.LGarXiv:2603.00610v32026
  39. Segment Anything Model for Medical Image Analysis: an Experimental Study

    Maciej A. Mazurowski, Haoyu Dong, Hanxue Gu +3

    cs.CVcs.AIcs.LGarXiv:2304.10517v32023
  40. Formalizing and Automating Fine-Grained Move Refactorings Across Methods

    Kota Yasuhara, Shinpei Hayashi

    cs.SEarXiv:2608.23377v12026
  41. Cooperative Non-Orthogonal Multiple Access with Simultaneous Wireless Information and Power Transfer

    Yuanwei Liu, Zhiguo Ding, Maged Elkashlan +1

    cs.ITarXiv:1511.02833v12015
  42. InfoPO: Information-Driven Policy Optimization for User-Centric Agents

    Fanqi Kong, Jiayi Zhang, Mingyi Deng +3

    cs.AIarXiv:2603.00656v22026
  43. A survey of graphical languages for monoidal categories

    Peter Selinger

    math.CTarXiv:0908.3347v12009
  44. LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation

    Jiaxiang Tang, Zhaoxi Chen, Xiaokang Chen +3

    cs.CVarXiv:2402.05054v12024
  45. Contrastive Clustering

    Yunfan Li, Peng Hu, Zitao Liu +3

    cs.LGcs.CVstat.MLarXiv:2009.09687v12020
  46. Future Optical Flow Prediction Improves Robot Control & Video Generation

    Kanchana Ranasinghe, Honglu Zhou, Yu Fang +7

    cs.CVarXiv:2601.10781v12026
  47. The All-Sky Automated Survey for Supernovae (ASAS-SN) Light Curve Server v1.0

    C. S. Kochanek, B. J. Shappee, K. Z. Stanek +9

    astro-ph.SRastro-ph.IMarXiv:1706.07060v12017
  48. Joint Optic Disc and Cup Segmentation Based on Multi-label Deep Network and Polar Transformation

    Huazhu Fu, Jun Cheng, Yanwu Xu +3

    cs.CVarXiv:1801.00926v32018
  49. WnW: Waxing-and-Waning KV Cache for Long-Form Speech LLMs

    Yiming Yao, Chenyang Lyu, Xuanfan Ni +4

    cs.CLcs.SDarXiv:2608.22704v12026
  50. ASVspoof 2019: Future Horizons in Spoofed and Fake Audio Detection

    Massimiliano Todisco, Xin Wang, Ville Vestman +7

    eess.AScs.CRcs.SDarXiv:1904.05441v22019
  51. VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding

    Jiapeng Shi, Junke Wang, Zuyao You +2

    cs.CVarXiv:2601.07290v12026
  52. Spend Search Where It Pays: Value-Guided Structured Sampling and Optimization for Generative Recommendation

    Jie Jiang, Yangru Huang, Zeyu Wang +4

    cs.AIcs.LGarXiv:2602.10699v22026
  53. KILT: a Benchmark for Knowledge Intensive Language Tasks

    Fabio Petroni, Aleksandra Piktus, Angela Fan +10

    cs.CLcs.AIcs.IRarXiv:2009.02252v42020
  54. Cyber-Security in Smart Grid: Survey and Challenges

    Zakaria El Mrabet, Hassan El Ghazi, Naima Kaabouch +1

    cs.CRcs.NIarXiv:1809.02609v12018
  55. Bilateral Multi-Perspective Matching for Natural Language Sentences

    Zhiguo Wang, Wael Hamza, Radu Florian

    cs.AIcs.CLarXiv:1702.03814v32017
  56. Definitional Sensitivity in Media Bias Detection: A Multi-Definition Dataset and Benchmark

    Martin Wessel, Timo Spinde, Jürgen Pfeffer +1

    cs.CLarXiv:2608.23095v12026
  57. Building machines that adapt and compute like brains

    Nikolaus Kriegeskorte, Robert M. Mok

    cs.AIq-bio.NCarXiv:1711.04203v12017
  58. High-Fidelity Audio Compression with Improved RVQGAN

    Rithesh Kumar, Prem Seetharaman, Alejandro Luebs +2

    cs.SDcs.LGeess.ASarXiv:2306.06546v22023
  59. SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generation

    Vaibhav Agrawal, Rishubh Parihar, Pradhaan Bhat +2

    cs.CVcs.AIarXiv:2602.23359v12026
  60. HydroShear: Hydroelastic Shear Simulation for Tactile Sim-to-Real Reinforcement Learning

    An Dang, Jayjun Lee, Mustafa Mukadam +4

    cs.ROcs.AIarXiv:2603.00446v12026