Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
25,921 to 25,980 of 61,350
Qwen3-Omni Technical Report
Jin Xu, Zhifang Guo, Hangrui Hu +35
cs.CLcs.AIcs.CVarXiv:2509.17765v12025The State of the Art in Enhancing Trust in Machine Learning Models with the Use of Visualizations
A. Chatzimparmpas, R. Martins, I. Jusufi +3
cs.LGcs.HCstat.MLarXiv:2212.11737v22022Artificial intelligence enabled radio propagation for communications-Part II: Scenario identification and channel modeling
Chen Huang, Ruisi He, Bo Ai +8
eess.SParXiv:2111.12228v12021Satellite Swarms for Direct-to-Cell Networks: A Distribution-Performance Trade-off Analysis
Xavier Artiga, Marius Caus, Ana I. Pérez-Neira +2
eess.SParXiv:2609.01380v12026Summaries:한국어A Benchmark for Lidar Sensors in Fog: Is Detection Breaking Down?
Mario Bijelic, Tobias Gruber, Werner Ritter
cs.CVarXiv:1912.03251v12019CrossFit: A Few-shot Learning Challenge for Cross-task Generalization in NLP
Qinyuan Ye, Bill Yuchen Lin, Xiang Ren
cs.CLcs.LGarXiv:2104.08835v22021Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Microsoft, :, Abdelrahman Abouelenin +73
cs.CLcs.AIcs.LGarXiv:2503.01743v22025Online Non-Monotone DR-Submodular Maximization Matching the Offline $0.401$ Factor
Vaneet Aggarwal, Yiyang Lu
cs.LGcs.AIcs.CCarXiv:2609.02145v12026Open-ended Learning in Symmetric Zero-sum Games
David Balduzzi, Marta Garnelo, Yoram Bachrach +4
cs.LGcs.GTcs.MAarXiv:1901.08106v22019IEEE 802.11be-Wi-Fi 7: New Challenges and Opportunities
Cailian Deng, Xuming Fang, Xiao Han +5
eess.SParXiv:2007.13401v32020Spot the conversation: speaker diarisation in the wild
Joon Son Chung, Jaesung Huh, Arsha Nagrani +2
cs.SDcs.CVeess.ASarXiv:2007.01216v32020Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
Siyan Zhao, Zhihui Xie, Mengchen Liu +4
cs.LGcs.CLarXiv:2601.18734v32026SNE-RoadSeg: Incorporating Surface Normal Information into Semantic Segmentation for Accurate Freespace Detection
Rui Fan, Hengli Wang, Peide Cai +1
cs.CVcs.ROeess.IVarXiv:2008.11351v12020InsightSeg: Reusing Correction Insights for Guideline-Consistent Segmentation
Vanshika Vats, Ashwani Rathee, James Davis
cs.CVcs.AIarXiv:2609.02002v12026Continuous 3D Perception Model with Persistent State
Qianqian Wang, Yifei Zhang, Aleksander Holynski +2
cs.CVarXiv:2501.12387v12025Weakly-Supervised Action Segmentation with Iterative Soft Boundary Assignment
Li Ding, Chenliang Xu
cs.CVarXiv:1803.10699v12018Video-R1: Reinforcing Video Reasoning in MLLMs
Kaituo Feng, Kaixiong Gong, Bohao Li +7
cs.CVarXiv:2503.21776v42025Sparse Instance Activation for Real-Time Instance Segmentation
Tianheng Cheng, Xinggang Wang, Shaoyu Chen +5
cs.CVarXiv:2203.12827v12022SCX Router: Streaming Zero-Shot Model Selection with a Decoder-KV Classifier and a Real-World Task Ontology
Ihor Stepanov, Aleksandr Smechov, Mykhailo Shtopko +2
cs.AIcs.CLarXiv:2609.02292v12026Linguistic Binding in Diffusion Models: Enhancing Attribute Correspondence through Attention Map Alignment
Royi Rassin, Eran Hirsch, Daniel Glickman +3
cs.CLcs.CVarXiv:2306.08877v32023VACE: All-in-One Video Creation and Editing
Zeyinzi Jiang, Zhen Han, Chaojie Mao +3
cs.CVarXiv:2503.07598v22025Codebook Agent: Amortized Topology Design for LLM Multi-Agent Systems
Jinxi Yu, Yubei Li, Eric Hanchen Jiang +6
cs.AIcs.LGcs.MAarXiv:2609.02264v12026Prediction Poisoning: Towards Defenses Against DNN Model Stealing Attacks
Tribhuvanesh Orekondy, Bernt Schiele, Mario Fritz
cs.LGcs.CRcs.CVarXiv:1906.10908v22019VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation
Xuan He, Dongfu Jiang, Ge Zhang +16
cs.CVcs.AIarXiv:2406.15252v32024BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents
Jason Wei, Zhiqing Sun, Spencer Papay +7
cs.CLarXiv:2504.12516v12025Propose to Learn, Learn to Propose: Evaluability-Aware Assistance under Bounded Rationality
Yifan Zhu, Sammie Katt, Samuel Kaski
cs.AIcs.HCcs.MAarXiv:2609.02242v12026SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Tianzhe Chu, Yuexiang Zhai, Jihan Yang +6
cs.AIcs.CVcs.LGarXiv:2501.17161v22025UI-TARS: Pioneering Automated GUI Interaction with Native Agents
Yujia Qin, Yining Ye, Junjie Fang +32
cs.AIcs.CLcs.CVarXiv:2501.12326v12025GoLLIE: Annotation Guidelines improve Zero-Shot Information-Extraction
Oscar Sainz, Iker García-Ferrero, Rodrigo Agerri +3
cs.CLarXiv:2310.03668v52023PhoenixNest-Video: Evidence-Grounded Multimodal Agent Framework for Automated Video Interview Assessment
Fan Yuxuan, Huang Miaojun, Zhang Haimei +2
cs.AIarXiv:2609.02231v12026Dream 7B: Diffusion Large Language Models
Jiacheng Ye, Zhihui Xie, Lin Zheng +5
cs.CLarXiv:2508.15487v12025RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
Tianxing Chen, Zanxin Chen, Baijun Chen +23
cs.ROcs.AIcs.CLarXiv:2506.18088v22025Motion-Aware Feature for Improved Video Anomaly Detection
Yi Zhu, Shawn Newsam
cs.CVcs.LGeess.IVarXiv:1907.10211v12019Deep Image Spatial Transformation for Person Image Generation
Yurui Ren, Xiaoming Yu, Junming Chen +2
cs.CVcs.AIarXiv:2003.00696v22020On Top-Down and Local Lower Bounds for $\mathrm{AC^0}$ Circuits
Gülce Kardeş, Benjamin Rossman
cs.CCarXiv:2609.01759v12026Mean Flows for One-step Generative Modeling
Zhengyang Geng, Mingyang Deng, Xingjian Bai +2
cs.LGcs.CVarXiv:2505.13447v12025Recent advances in opinion propagation dynamics: A 2020 Survey
Hossein Noorazar
physics.soc-phcs.SIarXiv:2004.05286v32020Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Prateek Chhikara, Dev Khant, Saket Aryan +2
cs.CLcs.AIarXiv:2504.19413v12025Interdependent networks with correlated degrees of mutually dependent nodes
Sergey V. Buldyrev, Nathaniel Shere, Gabriel A. Cwilich
cond-mat.dis-nncond-mat.stat-mecharXiv:1009.3183v12010Simulating electron energy loss spectroscopy with the MNPBEM toolbox
Ulrich Hohenester
cond-mat.mes-hallarXiv:1312.0748v12013Flow-GRPO: Training Flow Matching Models via Online RL
Jie Liu, Gongye Liu, Jiajun Liang +6
cs.CVcs.AIarXiv:2505.05470v52025V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Mido Assran, Adrien Bardes, David Fan +27
cs.AIcs.CVcs.LGarXiv:2506.09985v12025MemeCULT-1K: Benchmarking South Asian Cultural Context and Humor Understanding of Multimodal Models
Tawsif Tashwar Dipto, Mehedi Ahamed, Radib Bin Kabir +5
cs.CLcs.CVarXiv:2609.01772v12026Group Sequence Policy Optimization
Chujie Zheng, Shixuan Liu, Mingze Li +9
cs.LGcs.AIcs.CLarXiv:2507.18071v22025How will Language Modelers like ChatGPT Affect Occupations and Industries?
Ed Felten, Manav Raj, Robert Seamans
econ.GNarXiv:2303.01157v22023Categorizing Variants of Goodhart's Law
David Manheim, Scott Garrabrant
cs.AIq-fin.GNstat.MLarXiv:1803.04585v42018Blind Quality Assessment for in-the-Wild Images via Hierarchical Feature Fusion and Iterative Mixed Database Training
Wei Sun, Xiongkuo Min, Danyang Tu +2
cs.MMarXiv:2105.14550v32021Self-Training: A Survey
Massih-Reza Amini, Vasilii Feofanov, Loic Pauletto +3
cs.LGarXiv:2202.12040v62022Direct Satellite-to-Device Communications: From Cooperative Task Offloading to Non-Cooperative Access Monitoring
Sai Huang, Wanli Ni, Ke Lv +5
cs.ITeess.SParXiv:2609.02955v12026Towards Open-World Recommendation with Knowledge Augmentation from Large Language Models
Yunjia Xi, Weiwen Liu, Jianghao Lin +8
cs.IRarXiv:2306.10933v42023Identifying modular flows on multilayer networks reveals highly overlapping organization in social systems
Manlio De Domenico, Andrea Lancichinetti, Alex Arenas +1
physics.soc-phcs.SIarXiv:1408.2925v12014Human-AI Collaboration via Conditional Delegation: A Case Study of Content Moderation
Vivian Lai, Samuel Carton, Rajat Bhatnagar +3
cs.AIcs.HCcs.LGarXiv:2204.11788v12022Large-scale JPEG steganalysis using hybrid deep-learning framework
Jishen Zeng, Shunquan Tan, Bin Li +1
cs.MMarXiv:1611.03233v32016Occlusion-Robust Multimodal Emotion Recognition in VR via Fusion of Facial Images and EMG
Birgit Nierula, Karam Tomotaki-Dawoud, Mert Akguel +5
cs.CVcs.HCarXiv:2609.03569v12026Abnormal respiratory patterns classifier may contribute to large-scale screening of people infected with COVID-19 in an accurate and unobtrusive manner
Yunlu Wang, Menghan Hu, Qingli Li +3
cs.LGcs.CVeess.SParXiv:2002.05534v22020A Token-level Reference-free Hallucination Detection Benchmark for Free-form Text Generation
Tianyu Liu, Yizhe Zhang, Chris Brockett +4
cs.CLcs.AIarXiv:2104.08704v22021Practical Coreset Constructions for Machine Learning
Olivier Bachem, Mario Lucic, Andreas Krause
stat.MLarXiv:1703.06476v22017Multidimensional HLLE Riemann solver; Application to Euler and Magnetohydrodynamic Flows
Dinshaw S. Balsara
physics.comp-phphysics.flu-dynarXiv:0911.1613v12009CAPQ-FAST: Content-Adaptive Perceived Quality Assessment for Faster Audiovisual Playback
Jiarun Song, Yuxin Song, Fuzheng Yang +1
cs.MMarXiv:2609.03498v12026Conditional Generative Neural System for Probabilistic Trajectory Prediction
Jiachen Li, Hengbo Ma, Masayoshi Tomizuka
cs.CVcs.AIcs.LGarXiv:1905.01631v22019