Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
35,341 to 35,400 of 61,230
WHAM: Reconstructing World-grounded Humans with Accurate 3D Motion
Soyong Shin, Juyong Kim, Eni Halilaj +1
cs.CVarXiv:2312.07531v22023Support-set bottlenecks for video-text representation learning
Mandela Patrick, Po-Yao Huang, Yuki Asano +4
cs.CVarXiv:2010.02824v22020LogME: Practical Assessment of Pre-trained Models for Transfer Learning
Kaichao You, Yong Liu, Jianmin Wang +1
cs.LGcs.AIarXiv:2102.11005v32021Agent AI: Surveying the Horizons of Multimodal Interaction
Zane Durante, Qiuyuan Huang, Naoki Wake +11
cs.AIcs.HCcs.LGarXiv:2401.03568v22024Design of Task-Specific Optical Systems Using Broadband Diffractive Neural Networks
Yi Luo, Deniz Mengu, Nezih T. Yardimci +4
cs.NEphysics.comp-phphysics.opticsarXiv:1909.06553v12019Multi-Temporal Recurrent Neural Networks For Progressive Non-Uniform Single Image Deblurring With Incremental Temporal Training
Dongwon Park, Dong Un Kang, Jisoo Kim +1
eess.IVcs.CVarXiv:1911.07410v12019Large Language Models are Versatile Decomposers: Decompose Evidence and Questions for Table-based Reasoning
Yunhu Ye, Binyuan Hui, Min Yang +3
cs.CLarXiv:2301.13808v32023Exactly Diagonal Gram Matrices in Jacobi Weighted Histopolation
Allal Guessab, Federico Nudo, Stefano Serra-Capizzano
math.NAarXiv:2608.25714v12026LingxiDiagBench: A Multi-Agent Framework for Benchmarking LLMs in Chinese Psychiatric Consultation and Diagnosis
Shihao Xu, Tiancheng Zhou, Jiatong Ma +8
cs.AIcs.CLarXiv:2602.09379v32026TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions
Linli Yao, Yuancheng Wei, Yaojie Zhang +12
cs.CVarXiv:2602.08711v32026A Variational Perspective on Solving Inverse Problems with Diffusion Models
Morteza Mardani, Jiaming Song, Jan Kautz +1
cs.LGcs.CVmath.NAarXiv:2305.04391v22023Pathwise Test-Time Correction for Autoregressive Long Video Generation
Xunzhi Xiang, Zixuan Duan, Guiyu Zhang +7
cs.CVarXiv:2602.05871v22026PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards
Minh-Quan Le, Gaurav Mittal, Cheng Zhao +3
cs.CVarXiv:2602.01624v22026Towards Open Vocabulary Learning: A Survey
Jianzong Wu, Xiangtai Li, Shilin Xu +9
cs.CVcs.AIarXiv:2306.15880v42023Neural basis expansion analysis with exogenous variables: Forecasting electricity prices with NBEATSx
Kin G. Olivares, Cristian Challu, Grzegorz Marcjasz +2
cs.LGcs.AIstat.MLarXiv:2104.05522v62021PISA: Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers
Haopeng Li, Shitong Shao, Wenliang Zhong +4
cs.CVarXiv:2602.01077v22026A Tutorial on Fluid Antenna System for 6G Networks: Encompassing Communication Theory, Optimization Methods and Hardware Designs
Wee Kiat New, Kai-Kit Wong, Hao Xu +9
eess.SParXiv:2407.03449v22024How Proper Scoring Rules Shape LLM Forecasting
Benjamin Turtel, Paul Wilczewski, Kris Skotheim +2
cs.LGcs.AIarXiv:2608.28482v12026Observation of high-energy neutrinos from the Galactic plane
R. Abbasi, M. Ackermann, J. Adams +386
astro-ph.HEastro-ph.GAcs.LGarXiv:2307.04427v12023PGSR: Planar-based Gaussian Splatting for Efficient and High-Fidelity Surface Reconstruction
Danpeng Chen, Hai Li, Weicai Ye +7
cs.CVarXiv:2406.06521v220245G from Space: An Overview of 3GPP Non-Terrestrial Networks
Xingqin Lin, Stefan Rommer, Sebastian Euler +2
cs.NIeess.SParXiv:2103.09156v22021Show-1: Marrying Pixel and Latent Diffusion Models for Text-to-Video Generation
David Junhao Zhang, Jay Zhangjie Wu, Jia-Wei Liu +5
cs.CVarXiv:2309.15818v32023Navigate through Enigmatic Labyrinth A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future
Zheng Chu, Jingchang Chen, Qianglong Chen +7
cs.CLcs.AIarXiv:2309.15402v32023Defending against Backdoors in Federated Learning with Robust Learning Rate
Mustafa Safa Ozdayi, Murat Kantarcioglu, Yulia R. Gel
cs.LGcs.CRstat.MLarXiv:2007.03767v42020Stacked Intelligent Metasurfaces for Efficient Holographic MIMO Communications in 6G
Jiancheng An, Chao Xu, Derrick Wing Kwan Ng +4
cs.ITeess.SParXiv:2305.08079v12023Measuring the tendency of CNNs to Learn Surface Statistical Regularities
Jason Jo, Yoshua Bengio
cs.LGstat.MLarXiv:1711.11561v12017NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry
Samuel Xiao, Judy Song, Rory Hu +1
cs.CLcs.AIarXiv:2608.28481v12026Imagic: Text-Based Real Image Editing with Diffusion Models
Bahjat Kawar, Shiran Zada, Oran Lang +5
cs.CVarXiv:2210.09276v32022What Do We Want From Explainable Artificial Intelligence (XAI)? -- A Stakeholder Perspective on XAI and a Conceptual Model Guiding Interdisciplinary XAI Research
Markus Langer, Daniel Oster, Timo Speith +5
cs.AIcs.HCarXiv:2102.07817v12021More Diverse Means Better: Multimodal Deep Learning Meets Remote Sensing Imagery Classification
Danfeng Hong, Lianru Gao, Naoto Yokoya +4
cs.CVeess.IVarXiv:2008.05457v12020Reconfigurable Intelligent Surfaces: Principles and Opportunities
Yuanwei Liu, Xiao Liu, Xidong Mu +4
eess.SParXiv:2007.03435v32020Conv2Former: A Simple Transformer-Style ConvNet for Visual Recognition
Qibin Hou, Cheng-Ze Lu, Ming-Ming Cheng +1
cs.CVarXiv:2211.11943v12022On OTFS Modulation for High-Doppler Fading Channels
K. R. Murali, A. Chockalingam
cs.ITarXiv:1802.00929v12018The Hyper Suprime-Cam Software Pipeline
James Bosch, Robert Armstrong, Steven Bickerton +32
astro-ph.IMarXiv:1705.06766v12017On Unifying Multi-View Self-Representations for Clustering by Tensor Multi-Rank Minimization
Yuan Xie, Dacheng Tao, Wensheng Zhang +3
cs.CVarXiv:1610.07126v32016Random Feature Maps for Dot Product Kernels
Purushottam Kar, Harish Karnick
cs.LGcs.CGmath.FAarXiv:1201.6530v32012Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations
Ranjay Krishna, Yuke Zhu, Oliver Groth +9
cs.CVcs.AIarXiv:1602.07332v12016PTQ4ViT: Post-training quantization for vision transformers with twin uniform quantization
Zhihang Yuan, Chenhao Xue, Yiqi Chen +2
cs.CVarXiv:2111.12293v32021Dropout Sampling for Robust Object Detection in Open-Set Conditions
Dimity Miller, Lachlan Nicholson, Feras Dayoub +1
cs.CVarXiv:1710.06677v22017Improved Few-Shot Visual Classification
Peyman Bateni, Raghav Goyal, Vaden Masrani +2
cs.CVarXiv:1912.03432v32019Prioritized Training on Points that are Learnable, Worth Learning, and Not Yet Learnt
Sören Mindermann, Jan Brauner, Muhammed Razzak +8
cs.LGcs.AIcs.CLarXiv:2206.07137v32022Blog: Survey of Optimizers
Ruoran Xu
cs.LGcs.AIarXiv:2608.28557v12026Registration based Few-Shot Anomaly Detection
Chaoqin Huang, Haoyan Guan, Aofan Jiang +3
cs.CVarXiv:2207.07361v12022An Enclosed Mode Is a Gauge Choice: Topology Relative to Reach in Certified Code World Models
Javier Aguilar Martín
cs.LGcs.AIarXiv:2608.28541v12026Rewarded soups: towards Pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards
Alexandre Ramé, Guillaume Couairon, Mustafa Shukor +4
cs.LGcs.AIcs.CVarXiv:2306.04488v22023Implicit Reparameterization Gradients
Michael Figurnov, Shakir Mohamed, Andriy Mnih
cs.LGstat.MLarXiv:1805.08498v42018Asymmetric Student-Teacher Networks for Industrial Anomaly Detection
Marco Rudolph, Tom Wehrbein, Bodo Rosenhahn +1
cs.LGcs.AIcs.CVarXiv:2210.07829v22022Linearized two-layers neural networks in high dimension
Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz +1
math.STcs.LGarXiv:1904.12191v32019Texture Image Classification Using DWT AlexNet Feature Fusion and Deep Neural Networks
Arun D. Kulkarni
cs.CVcs.AIarXiv:2608.28524v12026Latent Embedding Feedback and Discriminative Features for Zero-Shot Classification
Sanath Narayan, Akshita Gupta, Fahad Shahbaz Khan +2
cs.CVarXiv:2003.07833v22020Large-Scale Methods for Distributionally Robust Optimization
Daniel Levy, Yair Carmon, John C. Duchi +1
math.OCcs.LGstat.MLarXiv:2010.05893v22020Learning Combinatorial Embedding Networks for Deep Graph Matching
Runzhong Wang, Junchi Yan, Xiaokang Yang
cs.CVarXiv:1904.00597v32019Conformal Uncertainty Quantification Guarantees for Neural Operators
Tom Stent, Nicolas Boullé
math.NAcs.AImath.PRarXiv:2608.28515v12026Learning to Answer Questions in Dynamic Audio-Visual Scenarios
Guangyao Li, Yake Wei, Yapeng Tian +3
cs.CVarXiv:2203.14072v22022ICE: Inter-instance Contrastive Encoding for Unsupervised Person Re-identification
Hao Chen, Benoit Lagadec, Francois Bremond
cs.CVarXiv:2103.16364v22021On the Maintenance and Co-evolution of Agent Plugins: An Empirical Study of Claude Code Plugin Marketplaces
Ahmed Hereiz, Yingzhe Lyu, Hao Li +2
cs.SEcs.AIarXiv:2608.28497v12026Communication and Control in Collaborative UAVs: Recent Advances and Future Trends
Shumaila Javaid, Nasir Saeed, Zakria Qadir +4
eess.SPcs.MAcs.ROarXiv:2302.12175v12023Resolving social dilemmas on evolving random networks
Attila Szolnoki, Matjaz Perc
physics.soc-phcond-mat.stat-mechq-bio.PEarXiv:0910.1905v12009Incorporating Nuisance Parameters in Likelihoods for Multisource Spectra
J. S. Conway
physics.data-anhep-exarXiv:1103.0354v12011Strong NP-Hardness of the Quantum Separability Problem
Sevag Gharibian
quant-pharXiv:0810.4507v52008