Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
46,021 to 46,080 of 61,114
CREPE: A Convolutional Representation for Pitch Estimation
Jong Wook Kim, Justin Salamon, Peter Li +1
eess.AScs.LGcs.SDarXiv:1802.06182v12018SPECTRA: Subspace-Preserving Embedding Calibration, Transport, and Replay for Fully Few-Shot Class-Incremental Audio Classification
Giries Abu Ayoub, Loay Mualem, Simon Korman
cs.SDcs.AIarXiv:2608.25054v12026Rethinking the Value of Labels for Improving Class-Imbalanced Learning
Yuzhe Yang, Zhi Xu
cs.LGcs.CVstat.MLarXiv:2006.07529v22020Learning to Walk via Deep Reinforcement Learning
Tuomas Haarnoja, Sehoon Ha, Aurick Zhou +3
cs.LGcs.AIcs.ROarXiv:1812.11103v32018A Study of BFLOAT16 for Deep Learning Training
Dhiraj Kalamkar, Dheevatsa Mudigere, Naveen Mellempudi +16
cs.LGstat.MLarXiv:1905.12322v32019Scissorhands: Exploiting the Persistence of Importance Hypothesis for LLM KV Cache Compression at Test Time
Zichang Liu, Aditya Desai, Fangshuo Liao +5
cs.LGcs.CLarXiv:2305.17118v22023Wireless Image Transmission Using Deep Source Channel Coding With Attention Modules
Jialong Xu, Bo Ai, Wei Chen +3
cs.ITarXiv:2012.00533v32020AppAgent: Multimodal Agents as Smartphone Users
Chi Zhang, Zhao Yang, Jiaxuan Liu +6
cs.CVarXiv:2312.13771v32023Rapidly Convergent Finite-Element Domain Decomposition Method With Two-Channel Transmission Conditions
Furkan Şık, Fernando L. Teixeira, Balasubramaniam Shanker
physics.comp-phmath.NAarXiv:2608.25041v12026Understanding Deep Networks via Extremal Perturbations and Smooth Masks
Ruth Fong, Mandela Patrick, Andrea Vedaldi
cs.CVcs.LGstat.MLarXiv:1910.08485v12019DiffusionDB: A Large-scale Prompt Gallery Dataset for Text-to-Image Generative Models
Zijie J. Wang, Evan Montoya, David Munechika +3
cs.CVcs.AIcs.HCarXiv:2210.14896v42022Subsampled Rényi Differential Privacy and Analytical Moments Accountant
Yu-Xiang Wang, Borja Balle, Shiva Kasiviswanathan
cs.LGcs.CRstat.MLarXiv:1808.00087v22018Empirical Risk Minimization under Fairness Constraints
Michele Donini, Luca Oneto, Shai Ben-David +2
stat.MLcs.LGarXiv:1802.08626v32018A Subcarrier-Aware Approach for Robust Respiratory Monitoring with Commodity Wi-Fi
Pei Tang, Yunpeng Ge, Ivan Wang-Hei Ho
eess.SParXiv:2608.25612v12026Dissonance Spectrum explicitly models perceptual frequency interactions for better music understanding
Tianle Wang, Xinyi Tong, Liangke Zhao +7
cs.SDcs.AIarXiv:2608.25621v12026Shap-E: Generating Conditional 3D Implicit Functions
Heewoo Jun, Alex Nichol
cs.CVcs.LGarXiv:2305.02463v12023InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents
Qiusi Zhan, Zhixiang Liang, Zifan Ying +1
cs.CLcs.CRarXiv:2403.02691v32024SimSwap: An Efficient Framework For High Fidelity Face Swapping
Renwang Chen, Xuanhong Chen, Bingbing Ni +1
cs.CVarXiv:2106.06340v12021Uncertainty Sets for Image Classifiers using Conformal Prediction
Anastasios Angelopoulos, Stephen Bates, Jitendra Malik +1
cs.CVmath.STstat.MLarXiv:2009.14193v52020Graph-Based Global Reasoning Networks
Yunpeng Chen, Marcus Rohrbach, Zhicheng Yan +3
cs.CVarXiv:1811.12814v12018Towards Fast, Accurate and Stable 3D Dense Face Alignment
Jianzhu Guo, Xiangyu Zhu, Yang Yang +3
cs.CVarXiv:2009.09960v22020Sparse Modeling for Image and Vision Processing
Julien Mairal, Francis Bach, Jean Ponce
cs.CVarXiv:1411.3230v22014TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale Up
Yifan Jiang, Shiyu Chang, Zhangyang Wang
cs.CVarXiv:2102.07074v42021Named Entity Recognition as Dependency Parsing
Juntao Yu, Bernd Bohnet, Massimo Poesio
cs.CLarXiv:2005.07150v32020Beyond Minimum Distance: The Optimal Leading Coefficient in the High-SNR Error-Probability Expansion for AWGN Spherical Codes
Nikola Zlatanov
cs.ITarXiv:2608.25805v12026TokenFlow: Consistent Diffusion Features for Consistent Video Editing
Michal Geyer, Omer Bar-Tal, Shai Bagon +1
cs.CVarXiv:2307.10373v32023DF-Net: Unsupervised Joint Learning of Depth and Flow using Cross-Task Consistency
Yuliang Zou, Zelun Luo, Jia-Bin Huang
cs.CVarXiv:1809.01649v12018On the Importance of Gradients for Detecting Distributional Shifts in the Wild
Rui Huang, Andrew Geng, Yixuan Li
cs.LGarXiv:2110.00218v22021LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models
Zhiqiang Hu, Lei Wang, Yihuai Lan +6
cs.CLarXiv:2304.01933v32023A Comprehensive Survey on Applications of Transformers for Deep Learning Tasks
Saidul Islam, Hanae Elmekki, Ahmed Elsebai +4
cs.LGcs.CLarXiv:2306.07303v12023FABRICA: Agentic CUDA-to-CSL Translation and Optimization for Wafer-Scale Systems
Yuebo Luo, Eliu Huerta, Venkatram Vishwanath +3
cs.CEarXiv:2608.25124v12026Demonstration of quantum volume 64 on a superconducting quantum computing system
Petar Jurcevic, Ali Javadi-Abhari, Lev S. Bishop +28
quant-pharXiv:2008.08571v22020Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction
Bowen Shi, Wei-Ning Hsu, Kushal Lakhotia +1
eess.AScs.CVcs.SDarXiv:2201.02184v22022Integer Quantization for Deep Learning Inference: Principles and Empirical Evaluation
Hao Wu, Patrick Judd, Xiaojie Zhang +2
cs.LGstat.MLarXiv:2004.09602v12020Spatial-Temporal Recurrent Neural Network for Emotion Recognition
Tong Zhang, Wenming Zheng, Zhen Cui +2
cs.CVarXiv:1705.04515v12017XGNN: Towards Model-Level Explanations of Graph Neural Networks
Hao Yuan, Jiliang Tang, Xia Hu +1
cs.LGstat.MLarXiv:2006.02587v12020LinkBERT: Pretraining Language Models with Document Links
Michihiro Yasunaga, Jure Leskovec, Percy Liang
cs.CLcs.LGarXiv:2203.15827v12022From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
Tianle Li, Wei-Lin Chiang, Evan Frick +5
cs.LGcs.AIcs.CLarXiv:2406.11939v22024A Comparative Evaluation of Digitization Pipelines for Historiographical Sources
Marina Gómez Rey, Patricia Callejo, Mario Muñoz-Organero +1
cs.DLarXiv:2608.24976v12026Scene Text Detection and Recognition: The Deep Learning Era
Shangbang Long, Xin He, Cong Yao
cs.CVarXiv:1811.04256v52018TransVG: End-to-End Visual Grounding with Transformers
Jiajun Deng, Zhengyuan Yang, Tianlang Chen +2
cs.CVarXiv:2104.08541v42021Revisiting Few-sample BERT Fine-tuning
Tianyi Zhang, Felix Wu, Arzoo Katiyar +2
cs.CLcs.LGarXiv:2006.05987v32020Illumination-aware Faster R-CNN for Robust Multispectral Pedestrian Detection
Chengyang Li, Dan Song, Ruofeng Tong +1
cs.CVarXiv:1803.05347v22018Fairness and Accountability Design Needs for Algorithmic Support in High-Stakes Public Sector Decision-Making
Michael Veale, Max Van Kleek, Reuben Binns
cs.CYcs.HCcs.LGarXiv:1802.01029v12018Compiling Spatial Certificates into Temporal Contracts for Latency-Aware Control
Avinash malik
eess.SYarXiv:2608.25228v12026Query Expansion Is More Than Generation: Improving Dense Retrieval through Better Integration
Siyuan Sun, Mihai Surdeanu
cs.IRarXiv:2608.25521v12026Time for a change: a tutorial for comparing multiple classifiers through Bayesian analysis
Alessio Benavoli, Giorgio Corani, Janez Demsar +1
stat.MLcs.LGarXiv:1606.04316v32016Detecting Language Model Attacks with Perplexity
Gabriel Alon, Michael Kamfonas
cs.CLcs.AIcs.CRarXiv:2308.14132v32023CoNet: Collaborative Cross Networks for Cross-Domain Recommendation
Guangneng Hu, Yu Zhang, Qiang Yang
cs.IRcs.AIcs.LGarXiv:1804.06769v32018Paging with Per-Replacement Maximum Delay
Tianhang Lu, Runtian Ren, Shengcai Liu
cs.DSarXiv:2608.25290v12026UniPC: A Unified Predictor-Corrector Framework for Fast Sampling of Diffusion Models
Wenliang Zhao, Lujia Bai, Yongming Rao +2
cs.LGcs.CVarXiv:2302.04867v42023Empirical Review of Automated Analysis Tools on 47,587 Ethereum Smart Contracts
Thomas Durieux, João F. Ferreira, Rui Abreu +1
cs.SEarXiv:1910.10601v22019ProphetNet: Predicting Future N-gram for Sequence-to-Sequence Pre-training
Weizhen Qi, Yu Yan, Yeyun Gong +5
cs.CLarXiv:2001.04063v32020Regressing Robust and Discriminative 3D Morphable Models with a very Deep Neural Network
Anh Tuan Tran, Tal Hassner, Iacopo Masi +1
cs.CVarXiv:1612.04904v12016Tabular Foundation Models for Multi-View Information Cascade Popularity Prediction
Wenting Zhu, Chenghua Gong, Sanchuan Guo +3
cs.SIarXiv:2608.25048v12026A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models
S. M Towhidul Islam Tonmoy, S M Mehedi Zaman, Vinija Jain +4
cs.CLarXiv:2401.01313v32024Deep learning for undersampled MRI reconstruction
Chang Min Hyun, Hwa Pyung Kim, Sung Min Lee +2
stat.MLcs.LGphysics.med-pharXiv:1709.02576v32017VL-Adapter: Parameter-Efficient Transfer Learning for Vision-and-Language Tasks
Yi-Lin Sung, Jaemin Cho, Mohit Bansal
cs.CVcs.AIcs.CLarXiv:2112.06825v22021Backpropagation Algorithms and Reservoir Computing in Recurrent Neural Networks for the Forecasting of Complex Spatiotemporal Dynamics
Pantelis R. Vlachas, Jaideep Pathak, Brian R. Hunt +4
eess.SPcs.LGphysics.flu-dynarXiv:1910.05266v22019DeepJSCC-f: Deep Joint Source-Channel Coding of Images with Feedback
David Burth Kurka, Deniz Gündüz
cs.ITcs.LGeess.IVarXiv:1911.11174v22019