Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

25,981 to 26,040 of 61,428

  1. DoFE: Domain-oriented Feature Embedding for Generalizable Fundus Image Segmentation on Unseen Datasets

    Shujun Wang, Lequan Yu, Kang Li +3

    cs.CVarXiv:2010.06208v12020
  2. The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models

    Ganqu Cui, Yuchen Zhang, Jiacheng Chen +14

    cs.LGcs.AIcs.CLarXiv:2505.22617v12025
  3. MedGemma Technical Report

    Andrew Sellergren, Sahar Kazemzadeh, Tiam Jaroensri +78

    cs.AIcs.CLcs.CVarXiv:2507.05201v42025
  4. Escaping Redundant Reasoning: Structure-Aware Search for Inference-Time LLMs

    Lu Cheng

    cs.AIarXiv:2609.00738v12026
  5. Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

    Wenxuan Huang, Bohan Jia, Zijie Zhai +7

    cs.CVcs.AIcs.CLarXiv:2503.06749v42025
  6. Grounded Human-Object Interaction Hotspots from Video

    Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman

    cs.CVarXiv:1812.04558v22018
  7. Simple linear attention language models balance the recall-throughput tradeoff

    Simran Arora, Sabri Eyuboglu, Michael Zhang +6

    cs.CLcs.LGarXiv:2402.18668v22024
  8. FAST: Efficient Action Tokenization for Vision-Language-Action Models

    Karl Pertsch, Kyle Stachowicz, Brian Ichter +6

    cs.ROcs.LGarXiv:2501.09747v12025
  9. FUSE: An Evaluating Framework for Dangerous Capabilities of LLMs

    Zhengyi Jin, Ru Zhang, Xiao Chen +5

    cs.AIarXiv:2609.02168v12026
  10. Social Learning and Distributed Hypothesis Testing

    Anusha Lalitha, Tara Javidi, Anand Sarwate

    math.STcs.ITmath.OCarXiv:1410.4307v52014
    Summaries:한국어
  11. Probabilistic Tools for the Analysis of Randomized Optimization Heuristics

    Benjamin Doerr

    cs.DScs.DMcs.NEarXiv:1801.06733v62018
  12. Why Do Multi-Agent LLM Systems Fail?

    Mert Cemri, Melissa Z. Pan, Shuyi Yang +10

    cs.AIarXiv:2503.13657v32025
  13. Comprehensive Graph-conditional Similarity Preserving Network for Unsupervised Cross-modal Hashing

    Jun Yu, Hao Zhou, Yibing Zhan +1

    cs.IRcs.CVarXiv:2012.13538v12020
  14. "It's a Fair Game", or Is It? Examining How Users Navigate Disclosure Risks and Benefits When Using LLM-Based Conversational Agents

    Zhiping Zhang, Michelle Jia, Hao-Ping Lee +5

    cs.HCcs.AIcs.CRarXiv:2309.11653v22023
  15. Understanding metric-related pitfalls in image analysis validation

    Annika Reinke, Minu D. Tizabi, Michael Baumgartner +75

    cs.CVarXiv:2302.01790v42023
  16. Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning

    Shenzhi Wang, Le Yu, Chang Gao +15

    cs.CLcs.AIcs.LGarXiv:2506.01939v22025
  17. Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space

    Chunyuan Li, Xiang Gao, Yuan Li +4

    cs.CLcs.LGstat.MLarXiv:2004.04092v42020
  18. MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

    Yubo Ma, Yuhang Zang, Liangyu Chen +13

    cs.CVcs.CLarXiv:2407.01523v32024
  19. Qwen3-Omni Technical Report

    Jin Xu, Zhifang Guo, Hangrui Hu +35

    cs.CLcs.AIcs.CVarXiv:2509.17765v12025
  20. The State of the Art in Enhancing Trust in Machine Learning Models with the Use of Visualizations

    A. Chatzimparmpas, R. Martins, I. Jusufi +3

    cs.LGcs.HCstat.MLarXiv:2212.11737v22022
  21. Artificial intelligence enabled radio propagation for communications-Part II: Scenario identification and channel modeling

    Chen Huang, Ruisi He, Bo Ai +8

    eess.SParXiv:2111.12228v12021
  22. Satellite Swarms for Direct-to-Cell Networks: A Distribution-Performance Trade-off Analysis

    Xavier Artiga, Marius Caus, Ana I. Pérez-Neira +2

    eess.SParXiv:2609.01380v12026
    Summaries:한국어
  23. A Benchmark for Lidar Sensors in Fog: Is Detection Breaking Down?

    Mario Bijelic, Tobias Gruber, Werner Ritter

    cs.CVarXiv:1912.03251v12019
  24. CrossFit: A Few-shot Learning Challenge for Cross-task Generalization in NLP

    Qinyuan Ye, Bill Yuchen Lin, Xiang Ren

    cs.CLcs.LGarXiv:2104.08835v22021
  25. Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

    Microsoft, :, Abdelrahman Abouelenin +73

    cs.CLcs.AIcs.LGarXiv:2503.01743v22025
  26. Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor

    Vaneet Aggarwal, Yiyang Lu

    cs.LGcs.AIcs.CCarXiv:2609.02145v12026
  27. Open-ended Learning in Symmetric Zero-sum Games

    David Balduzzi, Marta Garnelo, Yoram Bachrach +4

    cs.LGcs.GTcs.MAarXiv:1901.08106v22019
  28. IEEE 802.11be-Wi-Fi 7: New Challenges and Opportunities

    Cailian Deng, Xuming Fang, Xiao Han +5

    eess.SParXiv:2007.13401v32020
  29. Spot the conversation: speaker diarisation in the wild

    Joon Son Chung, Jaesung Huh, Arsha Nagrani +2

    cs.SDcs.CVeess.ASarXiv:2007.01216v32020
  30. Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

    Siyan Zhao, Zhihui Xie, Mengchen Liu +4

    cs.LGcs.CLarXiv:2601.18734v32026
  31. SNE-RoadSeg: Incorporating Surface Normal Information into Semantic Segmentation for Accurate Freespace Detection

    Rui Fan, Hengli Wang, Peide Cai +1

    cs.CVcs.ROeess.IVarXiv:2008.11351v12020
  32. InsightSeg: Reusing Correction Insights for Guideline-Consistent Segmentation

    Vanshika Vats, Ashwani Rathee, James Davis

    cs.CVcs.AIarXiv:2609.02002v12026
  33. Continuous 3D Perception Model with Persistent State

    Qianqian Wang, Yifei Zhang, Aleksander Holynski +2

    cs.CVarXiv:2501.12387v12025
  34. Weakly-Supervised Action Segmentation with Iterative Soft Boundary Assignment

    Li Ding, Chenliang Xu

    cs.CVarXiv:1803.10699v12018
  35. Video-R1: Reinforcing Video Reasoning in MLLMs

    Kaituo Feng, Kaixiong Gong, Bohao Li +7

    cs.CVarXiv:2503.21776v42025
  36. Sparse Instance Activation for Real-Time Instance Segmentation

    Tianheng Cheng, Xinggang Wang, Shaoyu Chen +5

    cs.CVarXiv:2203.12827v12022
  37. SCX Router: Streaming Zero-Shot Model Selection with a Decoder-KV Classifier and a Real-World Task Ontology

    Ihor Stepanov, Aleksandr Smechov, Mykhailo Shtopko +2

    cs.AIcs.CLarXiv:2609.02292v12026
  38. Linguistic Binding in Diffusion Models: Enhancing Attribute Correspondence through Attention Map Alignment

    Royi Rassin, Eran Hirsch, Daniel Glickman +3

    cs.CLcs.CVarXiv:2306.08877v32023
  39. VACE: All-in-One Video Creation and Editing

    Zeyinzi Jiang, Zhen Han, Chaojie Mao +3

    cs.CVarXiv:2503.07598v22025
  40. Codebook Agent: Amortized Topology Design for LLM Multi-Agent Systems

    Jinxi Yu, Yubei Li, Eric Hanchen Jiang +6

    cs.AIcs.LGcs.MAarXiv:2609.02264v12026
  41. Prediction Poisoning: Towards Defenses Against DNN Model Stealing Attacks

    Tribhuvanesh Orekondy, Bernt Schiele, Mario Fritz

    cs.LGcs.CRcs.CVarXiv:1906.10908v22019
  42. VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

    Xuan He, Dongfu Jiang, Ge Zhang +16

    cs.CVcs.AIarXiv:2406.15252v32024
  43. BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents

    Jason Wei, Zhiqing Sun, Spencer Papay +7

    cs.CLarXiv:2504.12516v12025
  44. Propose to Learn, Learn to Propose: Evaluability-Aware Assistance under Bounded Rationality

    Yifan Zhu, Sammie Katt, Samuel Kaski

    cs.AIcs.HCcs.MAarXiv:2609.02242v12026
  45. SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

    Tianzhe Chu, Yuexiang Zhai, Jihan Yang +6

    cs.AIcs.CVcs.LGarXiv:2501.17161v22025
  46. UI-TARS: Pioneering Automated GUI Interaction with Native Agents

    Yujia Qin, Yining Ye, Junjie Fang +32

    cs.AIcs.CLcs.CVarXiv:2501.12326v12025
  47. GoLLIE: Annotation Guidelines improve Zero-Shot Information-Extraction

    Oscar Sainz, Iker García-Ferrero, Rodrigo Agerri +3

    cs.CLarXiv:2310.03668v52023
  48. PhoenixNest-Video: Evidence-Grounded Multimodal Agent Framework for Automated Video Interview Assessment

    Fan Yuxuan, Huang Miaojun, Zhang Haimei +2

    cs.AIarXiv:2609.02231v12026
  49. Dream 7B: Diffusion Large Language Models

    Jiacheng Ye, Zhihui Xie, Lin Zheng +5

    cs.CLarXiv:2508.15487v12025
  50. RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation

    Tianxing Chen, Zanxin Chen, Baijun Chen +23

    cs.ROcs.AIcs.CLarXiv:2506.18088v22025
  51. Motion-Aware Feature for Improved Video Anomaly Detection

    Yi Zhu, Shawn Newsam

    cs.CVcs.LGeess.IVarXiv:1907.10211v12019
  52. Deep Image Spatial Transformation for Person Image Generation

    Yurui Ren, Xiaoming Yu, Junming Chen +2

    cs.CVcs.AIarXiv:2003.00696v22020
  53. On Top-Down and Local Lower Bounds for $\mathrm{AC^0}$ Circuits

    Gülce Kardeş, Benjamin Rossman

    cs.CCarXiv:2609.01759v12026
  54. Mean Flows for One-step Generative Modeling

    Zhengyang Geng, Mingyang Deng, Xingjian Bai +2

    cs.LGcs.CVarXiv:2505.13447v12025
  55. Recent advances in opinion propagation dynamics: A 2020 Survey

    Hossein Noorazar

    physics.soc-phcs.SIarXiv:2004.05286v32020
  56. Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

    Prateek Chhikara, Dev Khant, Saket Aryan +2

    cs.CLcs.AIarXiv:2504.19413v12025
  57. Interdependent networks with correlated degrees of mutually dependent nodes

    Sergey V. Buldyrev, Nathaniel Shere, Gabriel A. Cwilich

    cond-mat.dis-nncond-mat.stat-mecharXiv:1009.3183v12010
  58. Simulating electron energy loss spectroscopy with the MNPBEM toolbox

    Ulrich Hohenester

    cond-mat.mes-hallarXiv:1312.0748v12013
  59. Flow-GRPO: Training Flow Matching Models via Online RL

    Jie Liu, Gongye Liu, Jiajun Liang +6

    cs.CVcs.AIarXiv:2505.05470v52025
  60. V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

    Mido Assran, Adrien Bardes, David Fan +27

    cs.AIcs.CVcs.LGarXiv:2506.09985v12025