Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
56,161 to 56,220 of 61,224
StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, Siva Reddy
cs.CLcs.AIcs.CYarXiv:2004.09456v12020Position: LLM Inference Should Be Evaluated as Energy-to-Token Production
Xiang Liu, Shimiao Yuan, Zhenheng Tang +5
cs.CEcs.DCarXiv:2605.11733v12026Concept Bottleneck Models
Pang Wei Koh, Thao Nguyen, Yew Siang Tang +4
cs.LGstat.MLarXiv:2007.04612v32020Orthrus: Memory-Efficient Parallel Token Generation via Dual-View Diffusion
Chien Van Nguyen, Chaitra Hegde, Van Cuong Pham +3
cs.LGcs.AIarXiv:2605.12825v22026Closing the AI Accountability Gap: Defining an End-to-End Framework for Internal Algorithmic Auditing
Inioluwa Deborah Raji, Andrew Smart, Rebecca N. White +6
cs.CYarXiv:2001.00973v12020The DAWN of World-Action Interactive Models
Hongbo Lu, Liang Yao, Chenghao He +6
cs.CVarXiv:2605.11550v12026Ditto: Fair and Robust Federated Learning Through Personalization
Tian Li, Shengyuan Hu, Ahmad Beirami +1
cs.LGstat.MLarXiv:2012.04221v32020Scaled-YOLOv4: Scaling Cross Stage Partial Network
Chien-Yao Wang, Alexey Bochkovskiy, Hong-Yuan Mark Liao
cs.CVcs.LGarXiv:2011.08036v22020Efficient Object Localization Using Convolutional Networks
Jonathan Tompson, Ross Goroshin, Arjun Jain +2
cs.CVarXiv:1411.4280v32014BloombergGPT: A Large Language Model for Finance
Shijie Wu, Ozan Irsoy, Steven Lu +6
cs.LGcs.AIcs.CLarXiv:2303.17564v32023DocAtlas: Multilingual Document Understanding Across 80+ Languages
Ahmed Heakl, Youssef Mohamed, Abdullah Sohail +6
cs.CLcs.CVcs.LGarXiv:2605.12623v22026MedMNIST v2 -- A large-scale lightweight benchmark for 2D and 3D biomedical image classification
Jiancheng Yang, Rui Shi, Donglai Wei +5
cs.CVcs.AIcs.LGarXiv:2110.14795v220213D MRI brain tumor segmentation using autoencoder regularization
Andriy Myronenko
cs.CVq-bio.NCarXiv:1810.11654v32018Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions
Wenhai Wang, Enze Xie, Xiang Li +6
cs.CVarXiv:2102.12122v22021MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities
Mingqiao Ye, Zhaochong An, Zhitong Gao +11
cs.CVcs.AIcs.LGarXiv:2607.25948v12026On the Road to 6G: Visions, Requirements, Key Technologies and Testbeds
Cheng-Xiang Wang, Xiaohu You, Xiqi Gao +16
eess.SParXiv:2302.14536v12023Statistical physics of human cooperation
Matjaz Perc, Jillian J. Jordan, David G. Rand +3
physics.soc-phcond-mat.stat-mechcs.SIarXiv:1705.07161v12017Proximal alternating minimization and projection methods for nonconvex problems. An approach based on the Kurdyka-Lojasiewicz inequality
Hedy Attouch, Jerome Bolte, Patrick Redont +1
math.OCarXiv:0801.1780v32008DeepMind Control Suite
Yuval Tassa, Yotam Doron, Alistair Muldal +9
cs.AIarXiv:1801.00690v12018Frequency Bias and OOD Generalization in Neural Operators under a Variable-Coefficient Wave Equation
Runlong Xie, An Luo
cs.LGarXiv:2605.12997v12026Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds
Prateek Biswas, Dhaval Patel, Vedant Khandelwal +2
cs.IRcs.LGcs.PLarXiv:2605.18827v12026PersonalAI 2.0: Enhancing knowledge graph traversal/retrieval with planning mechanism for Personalized LLM Agents
Mikhail Menschikov, Matvey Iskornev, Alexander Kharitonov +8
cs.CLarXiv:2605.13481v12026Toward Multimodal Image-to-Image Translation
Jun-Yan Zhu, Richard Zhang, Deepak Pathak +4
cs.CVcs.GRstat.MLarXiv:1711.11586v42017Random Numbers Certified by Bell's Theorem
S. Pironio, A. Acin, S. Massar +8
quant-pharXiv:0911.3427v32009CTRL: A Conditional Transformer Language Model for Controllable Generation
Nitish Shirish Keskar, Bryan McCann, Lav R. Varshney +2
cs.CLarXiv:1909.05858v22019An open access repository of images on plant health to enable the development of mobile disease diagnostics
David. P. Hughes, Marcel Salathe
cs.CYarXiv:1511.08060v22015Error mitigation for short-depth quantum circuits
Kristan Temme, Sergey Bravyi, Jay M. Gambetta
quant-phcond-mat.otherarXiv:1612.02058v32016SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot
Elias Frantar, Dan Alistarh
cs.LGarXiv:2301.00774v32023Financial Time Series Forecasting with Deep Learning : A Systematic Literature Review: 2005-2019
Omer Berat Sezer, Mehmet Ugur Gudelek, Ahmet Murat Ozbayoglu
cs.LGq-fin.CPstat.MLarXiv:1911.13288v12019SPIN: Structural LLM Planning via Iterative Navigation for Industrial Tasks
Yusuke Ozaki, Dhaval Patel
cs.AIarXiv:2605.14051v12026Large-scale Point Cloud Semantic Segmentation with Superpoint Graphs
Loic Landrieu, Martin Simonovsky
cs.CVcs.LGcs.NEarXiv:1711.09869v22017Qwen2.5-Coder Technical Report
Binyuan Hui, Jian Yang, Zeyu Cui +21
cs.CLarXiv:2409.12186v32024Diffusion-LM Improves Controllable Text Generation
Xiang Lisa Li, John Thickstun, Ishaan Gulrajani +2
cs.CLcs.AIcs.LGarXiv:2205.14217v12022PRISM: Prior Rectification and Uncertainty-Aware Structure Modeling for Diffusion-Based Text Image Super-Resolution
Zihang Xu, Xiaoyang Liu, Zheng Chen +2
cs.CVarXiv:2605.13027v12026Do Vision Transformers See Like Convolutional Neural Networks?
Maithra Raghu, Thomas Unterthiner, Simon Kornblith +2
cs.CVcs.AIcs.LGarXiv:2108.08810v22021FBNet: Hardware-Aware Efficient ConvNet Design via Differentiable Neural Architecture Search
Bichen Wu, Xiaoliang Dai, Peizhao Zhang +7
cs.CVarXiv:1812.03443v32018StyleCLIP: Text-Driven Manipulation of StyleGAN Imagery
Or Patashnik, Zongze Wu, Eli Shechtman +2
cs.CVcs.CLcs.GRarXiv:2103.17249v12021LoREnc: Low-Rank Encryption for Securing Foundation Models and LoRA Adapters
Beomjin Ahn, Jungmin Kwon, Chanyong Jung +1
cs.CRcs.CVcs.LGarXiv:2605.13163v12026Towards Recursive Self-Evolving Agentic Literature Retrieval
Yuwen Du, Tian Jin, Jing Kang +8
cs.IRarXiv:2605.14306v32026AraBERT: Transformer-based Model for Arabic Language Understanding
Wissam Antoun, Fady Baly, Hazem Hajj
cs.CLarXiv:2003.00104v42020Training Large Language Models to Predict Clinical Events
Benjamin Turtel, Paul Wilczewski, Kris Skotheim
cs.LGcs.AIcs.CLarXiv:2605.12817v12026GPT Understands, Too
Xiao Liu, Yanan Zheng, Zhengxiao Du +4
cs.CLcs.LGarXiv:2103.10385v22021Train faster, generalize better: Stability of stochastic gradient descent
Moritz Hardt, Benjamin Recht, Yoram Singer
cs.LGmath.OCstat.MLarXiv:1509.01240v22015FlowCompile: An Optimizing Compiler for Structured LLM Workflows
Junyan Li, Zhang-Wei Hong, Maohao Shen +2
cs.CLarXiv:2605.13647v12026Unifying Visual-Semantic Embeddings with Multimodal Neural Language Models
Ryan Kiros, Ruslan Salakhutdinov, Richard S. Zemel
cs.LGcs.CLcs.CVarXiv:1411.2539v12014Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition
Kush Juvekar, Kavya Manohar, Aditya Srinivas Menon +2
cs.CLcs.AIarXiv:2605.13087v22026Context Training with Active Information Seeking
Zeyu Huang, Adhiguna Kuncoro, Qixuan Feng +4
cs.CLcs.AIarXiv:2605.13050v22026The Secret Sharer: Evaluating and Testing Unintended Memorization in Neural Networks
Nicholas Carlini, Chang Liu, Úlfar Erlingsson +2
cs.LGcs.AIcs.CRarXiv:1802.08232v32018RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior Imitation
Chengzhi Shen, Weixiang Shen, Tobias Susetzky +8
cs.AIcs.CLcs.LGarXiv:2605.13542v12026Lifelong Learning with Dynamically Expandable Networks
Jaehong Yoon, Eunho Yang, Jeongtae Lee +1
cs.LGarXiv:1708.01547v112017MAP: A Map-then-Act Paradigm for Long-Horizon Interactive Agent Reasoning
Yuxin Liu, Ziang Ye, Yueqing Sun +6
cs.AIarXiv:2605.13037v12026Deep Convolutional Neural Networks and Data Augmentation for Environmental Sound Classification
Justin Salamon, Juan Pablo Bello
cs.SDcs.CVcs.LGarXiv:1608.04363v22016Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Wei-Lin Chiang, Lianmin Zheng, Ying Sheng +8
cs.AIcs.CLarXiv:2403.04132v12024BOOKMARKS: Efficient Active Storyline Memory for Role-playing
Letian Peng, Ziche Liu, Yiming Huang +4
cs.CLarXiv:2605.14169v12026Stereo Matching by Training a Convolutional Neural Network to Compare Image Patches
Jure Žbontar, Yann LeCun
cs.CVcs.LGcs.NEarXiv:1510.05970v22015A Unified Framework for High-Dimensional Analysis of M-Estimators with Decomposable Regularizers
Sahand N. Negahban, Pradeep Ravikumar, Martin J. Wainwright +1
math.STcs.ITstat.MEarXiv:1010.2731v32010Bag of Tricks and A Strong Baseline for Deep Person Re-identification
Hao Luo, Youzhi Gu, Xingyu Liao +2
cs.CVarXiv:1903.07071v32019Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
Shan Yang
cs.CLarXiv:2605.14040v12026ELATE: An open-source online application for analysis and visualization of elastic tensors
Romain Gaillac, Pluton Pullumbi, François-Xavier Coudert
cond-mat.mtrl-sciphysics.comp-pharXiv:1602.06175v22016Learning POMDP World Models from Observations with Language-Model Priors
Valentin Six, Frederik Panse, Mathis Fajeau +7
cs.LGarXiv:2605.13740v12026