Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
55,561 to 55,620 of 61,271
Secure Transmission with Multiple Antennas: The MISOME Wiretap Channel
Ashish Khisti, Gregory Wornell
cs.ITarXiv:0708.4219v12007CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval
Huaishao Luo, Lei Ji, Ming Zhong +4
cs.CVarXiv:2104.08860v22021GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction
Hongxin Li, Yuntao Chen, Zhaoxiang Zhang
cs.CVarXiv:2604.23941v12026S^3-Rec: Self-Supervised Learning for Sequential Recommendation with Mutual Information Maximization
Kun Zhou, Hui Wang, Wayne Xin Zhao +5
cs.IRcs.LGarXiv:2008.07873v12020StackGAN++: Realistic Image Synthesis with Stacked Generative Adversarial Networks
Han Zhang, Tao Xu, Hongsheng Li +4
cs.CVcs.AIstat.MLarXiv:1710.10916v32017Composition-based Multi-Relational Graph Convolutional Networks
Shikhar Vashishth, Soumya Sanyal, Vikram Nitin +1
cs.LGstat.MLarXiv:1911.03082v22019ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Lin Chen, Jinsong Li, Xiaoyi Dong +5
cs.CVarXiv:2311.12793v22023Enhanced LSTM for Natural Language Inference
Qian Chen, Xiaodan Zhu, Zhenhua Ling +3
cs.CLarXiv:1609.06038v32016BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate
Arnon Mazza, Elad Levi
cs.CLcs.AIcs.LGarXiv:2604.25203v12026Symmetric Cross Entropy for Robust Learning with Noisy Labels
Yisen Wang, Xingjun Ma, Zaiyi Chen +3
cs.LGcs.CVstat.MLarXiv:1908.06112v12019Deep Learning for Medical Image Processing: Overview, Challenges and Future
Muhammad Imran Razzak, Saeeda Naz, Ahmad Zaib
cs.CVarXiv:1704.06825v12017Better Models, Faster Training: Sigmoid Attention for single-cell Foundation Models
Vijay Sadashivaiah, Georgios Dasoulas, Judith Mueller +1
cs.LGq-bio.QMarXiv:2604.27124v12026Gemma: Open Models Based on Gemini Research and Technology
Gemma Team, Thomas Mesnard, Cassidy Hardin +105
cs.CLcs.AIarXiv:2403.08295v42024MDETR -- Modulated Detection for End-to-End Multi-Modal Understanding
Aishwarya Kamath, Mannat Singh, Yann LeCun +3
cs.CVcs.CLcs.LGarXiv:2104.12763v22021FDA: Fourier Domain Adaptation for Semantic Segmentation
Yanchao Yang, Stefano Soatto
cs.CVarXiv:2004.05498v12020Scaling Laws for Reward Model Overoptimization
Leo Gao, John Schulman, Jacob Hilton
cs.LGstat.MLarXiv:2210.10760v12022How to Explain Individual Classification Decisions
David Baehrens, Timon Schroeter, Stefan Harmeling +3
stat.MLcs.LGarXiv:0912.1128v12009RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning
Yan Duan, John Schulman, Xi Chen +3
cs.AIcs.LGcs.NEarXiv:1611.02779v22016Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models
Xingwei Tan, Marco Valentino, Mahmud Elahi Akhter +3
cs.CLcs.AIarXiv:2604.27251v22026State-of-the-art Speech Recognition With Sequence-to-Sequence Models
Chung-Cheng Chiu, Tara N. Sainath, Yonghui Wu +11
cs.CLcs.SDeess.ASarXiv:1712.01769v62017Voxel R-CNN: Towards High Performance Voxel-based 3D Object Detection
Jiajun Deng, Shaoshuai Shi, Peiwei Li +3
cs.CVarXiv:2012.15712v22020Latent Retrieval for Weakly Supervised Open Domain Question Answering
Kenton Lee, Ming-Wei Chang, Kristina Toutanova
cs.CLarXiv:1906.00300v32019Large Language Models are not Fair Evaluators
Peiyi Wang, Lei Li, Liang Chen +7
cs.CLcs.AIcs.IRarXiv:2305.17926v22023Large Language Models Explore by Latent Distilling
Yuanhao Zeng, Ao Lu, Lufei Li +3
cs.CLcs.AIcs.LGarXiv:2604.24927v22026Vital nodes identification in complex networks
Linyuan Lü, Duanbing Chen, Xiao-Long Ren +3
physics.soc-phcs.SIarXiv:1607.01134v12016Safety Drift After Fine-Tuning: Evidence from High-Stakes Domains
Emaan Bilal Khan, Amy Winecoff, Miranda Bogen +1
cs.CYcs.SEarXiv:2604.24902v12026Deep Learning: A Critical Appraisal
Gary Marcus
cs.AIcs.LGstat.MLarXiv:1801.00631v12018RWKV: Reinventing RNNs for the Transformer Era
Bo Peng, Eric Alcaide, Quentin Anthony +31
cs.CLcs.AIarXiv:2305.13048v22023X2SAM: Any Segmentation in Images and Videos
Hao Wang, Limeng Qiao, Chi Zhang +4
cs.CVcs.AIarXiv:2605.00891v12026Accelerating 3D Deep Learning with PyTorch3D
Nikhila Ravi, Jeremy Reizenstein, David Novotny +4
cs.CVcs.GRcs.LGarXiv:2007.08501v12020FcaNet: Frequency Channel Attention Networks
Zequn Qin, Pengyi Zhang, Fei Wu +1
cs.CVarXiv:2012.11879v42020Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models
Jiayi Guo, Linqing Wang, Jiangshan Wang +6
cs.CVarXiv:2604.25636v12026Importance Estimation for Neural Network Pruning
Pavlo Molchanov, Arun Mallya, Stephen Tyree +2
cs.LGcs.CVstat.MLarXiv:1906.10771v12019RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments
Zaid Nasser, Mikhail Iumanov, Tianhao Li +3
cs.CVarXiv:2604.26067v12026Prior-Aligned Data Cleaning for Tabular Foundation Models
Laure Berti-Equille
cs.LGcs.DBarXiv:2604.25154v12026Learning to Reconstruct 3D Human Pose and Shape via Model-fitting in the Loop
Nikos Kolotouros, Georgios Pavlakos, Michael J. Black +1
cs.CVarXiv:1909.12828v12019RULER: What's the Real Context Size of Your Long-Context Language Models?
Cheng-Ping Hsieh, Simeng Sun, Samuel Kriman +5
cs.CLarXiv:2404.06654v32024Pseudo-LiDAR from Visual Depth Estimation: Bridging the Gap in 3D Object Detection for Autonomous Driving
Yan Wang, Wei-Lun Chao, Divyansh Garg +3
cs.CVarXiv:1812.07179v62018LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale
Tim Dettmers, Mike Lewis, Younes Belkada +1
cs.LGcs.AIarXiv:2208.07339v22022FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
Yanting Wang, Chenlong Yin, Ying Chen +1
cs.CRarXiv:2604.28157v12026Instruction-Guided Poetry Generation in Arabic and Its Dialects
Abdelrahman Sadallah, Kareem Elozeiri, Mervat Abassy +5
cs.CLcs.AIarXiv:2604.27766v12026Agentic Fusion of Large Atomic and Language Models to Accelerate Superconductor Discovery
Mingze Li, Yu Rong, Songyou Li +16
cs.LGcond-mat.mtrl-sciarXiv:2604.23758v32026Exploring the Limits of Language Modeling
Rafal Jozefowicz, Oriol Vinyals, Mike Schuster +2
cs.CLarXiv:1602.02410v22016FASH-iCNN: Making Editorial Fashion Identity Inspectable Through Multimodal CNN Probing
Morayo Danielle Adeyemi, Ryan A. Rossi, Franck Dernoncourt
cs.CVcs.HCcs.IRarXiv:2604.26186v12026Structural-RNN: Deep Learning on Spatio-Temporal Graphs
Ashesh Jain, Amir R. Zamir, Silvio Savarese +1
cs.CVcs.LGcs.NEarXiv:1511.05298v32015Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar +2
cs.CLcs.AIarXiv:1804.06876v12018Relief-Based Feature Selection: Introduction and Review
Ryan J. Urbanowicz, Melissa Meeker, William LaCava +2
cs.DScs.LGstat.MLarXiv:1711.08421v22017Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models
Gongbo Zhang, Wen Wang, Ye Tian +1
cs.CLcs.AIcs.LGarXiv:2604.26951v12026ML-Leaks: Model and Data Independent Membership Inference Attacks and Defenses on Machine Learning Models
Ahmed Salem, Yang Zhang, Mathias Humbert +3
cs.CRcs.AIcs.LGarXiv:1806.01246v22018DKN: Deep Knowledge-Aware Network for News Recommendation
Hongwei Wang, Fuzheng Zhang, Xing Xie +1
stat.MLcs.LGarXiv:1801.08284v22018From Coarse to Fine: Robust Hierarchical Localization at Large Scale
Paul-Edouard Sarlin, Cesar Cadena, Roland Siegwart +1
cs.CVarXiv:1812.03506v22018Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation
Jay Zhangjie Wu, Yixiao Ge, Xintao Wang +7
cs.CVarXiv:2212.11565v22022Captum: A unified and generic model interpretability library for PyTorch
Narine Kokhlikyan, Vivek Miglani, Miguel Martin +8
cs.LGcs.AIstat.MLarXiv:2009.07896v12020Action Recognition with Trajectory-Pooled Deep-Convolutional Descriptors
Limin Wang, Yu Qiao, Xiaoou Tang
cs.CVarXiv:1505.04868v12015DRAEM -- A discriminatively trained reconstruction embedding for surface anomaly detection
Vitjan Zavrtanik, Matej Kristan, Danijel Skočaj
cs.CVarXiv:2108.07610v22021Image De-raining Using a Conditional Generative Adversarial Network
He Zhang, Vishwanath Sindagi, Vishal M. Patel
cs.CVarXiv:1701.05957v42017Manopt, a Matlab toolbox for optimization on manifolds
Nicolas Boumal, Bamdev Mishra, P. -A. Absil +1
cs.MScs.LGmath.OCarXiv:1308.5200v12013Deep Learning for IoT Big Data and Streaming Analytics: A Survey
Mehdi Mohammadi, Ala Al-Fuqaha, Sameh Sorour +1
cs.NIcs.DBcs.LGarXiv:1712.04301v22017PLUTO: a Numerical Code for Computational Astrophysics
A. Mignone, G. Bodo, S. Massaglia +4
astro-pharXiv:astro-ph/0701854v22007Oriented R-CNN for Object Detection
Xingxing Xie, Gong Cheng, Jiabao Wang +2
cs.CVarXiv:2108.05699v12021