Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,561 to 4,620 of 61,350
ASR is all you need: cross-modal distillation for lip reading
Triantafyllos Afouras, Joon Son Chung, Andrew Zisserman
cs.CVcs.SDeess.ASarXiv:1911.12747v22019Component-Aware Differential Privacy for Federated Multilingual Speech-LLMs
Jordi Luque, Fernando López, Aleix Sant
cs.CLarXiv:2609.11762v12026Shape-Aware Organ Segmentation by Predicting Signed Distance Maps
Yuan Xue, Hui Tang, Zhi Qiao +6
cs.CVarXiv:1912.03849v12019Reconfigurable Intelligent Surfaces 2.0: Beyond Diagonal Phase Shift Matrices
Hongyu Li, Shanpu Shen, Matteo Nerini +1
eess.SPcs.ITarXiv:2301.03288v32023On the Use of BERT for Automated Essay Scoring: Joint Learning of Multi-Scale Essay Representation
Yongjie Wang, Chuan Wang, Ruobing Li +1
cs.CLcs.AIarXiv:2205.03835v22022SeaLLMs -- Large Language Models for Southeast Asia
Xuan-Phi Nguyen, Wenxuan Zhang, Xin Li +14
cs.CLarXiv:2312.00738v22023RAG-Safety-Bench: Reliable Evaluation of Retrieval-Augmented LLM Safety
Adithiyan Rajan Indira Saravanan, Kathleen C. Fraser
cs.CLcs.IRarXiv:2609.11758v12026FeynCalc 10: Do multiloop integrals dream of computer codes?
Vladyslav Shtabovenko, Rolf Mertig, Frederik Orellana
hep-phhep-tharXiv:2312.14089v22023Automated Variational Inference in Probabilistic Programming
David Wingate, Theophane Weber
stat.MLcs.AIcs.LGarXiv:1301.1299v12013Recurrent Convolutional Neural Networks for Scene Parsing
Pedro H. O. Pinheiro, Ronan Collobert
cs.CVarXiv:1306.2795v12013LOCUS: Task-Aware Low-Rank Post-Training for Token-Efficient Language Generation
Dongfang Zhao
cs.CLcs.AIcs.LGarXiv:2609.11739v12026Catalyst Acceleration for First-order Convex Optimization: from Theory to Practice
Hongzhou Lin, Julien Mairal, Zaid Harchaoui
stat.MLmath.OCarXiv:1712.05654v22017Obstacle Tower: A Generalization Challenge in Vision, Control, and Planning
Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges +6
cs.AIcs.LGarXiv:1902.01378v22019Manufacturable 300mm platform solution for Field-Free Switching SOT-MRAM
K. Garello, F. Yasin, H. Hody +14
physics.app-pharXiv:1907.08012v22019Towards neural networks that provably know when they don't know
Alexander Meinke, Matthias Hein
cs.LGcs.CVstat.MLarXiv:1909.12180v22019The Eloquence submission for Task 2 of the Interspeech 2026 MLC-SLM challenge
Jordi Luque, Lorenzo Concina, Marco Matassoni +2
cs.CLarXiv:2609.11724v12026IndoNLG: Benchmark and Resources for Evaluating Indonesian Natural Language Generation
Samuel Cahyawijaya, Genta Indra Winata, Bryan Wilie +9
cs.CLarXiv:2104.08200v32021Global Optimality in Low-rank Matrix Optimization
Zhihui Zhu, Qiuwei Li, Gongguo Tang +1
cs.ITmath.OCarXiv:1702.07945v32017Negative Self-Distillation: Learning to Reason by Avoiding Flaws
Rongcan Pei, Zhepei Wei, Shuyao Xu +3
cs.CLcs.LGarXiv:2609.11699v12026Global Encoding for Abstractive Summarization
Junyang Lin, Xu Sun, Shuming Ma +1
cs.CLcs.AIcs.LGarXiv:1805.03989v22018SDCNet: Video Prediction Using Spatially-Displaced Convolution
Fitsum A. Reda, Guilin Liu, Kevin J. Shih +5
cs.CVarXiv:1811.00684v22018Complex-Text Robustness Evaluation and Failure Diagnosis for Low-Resource Multilingual Text-to-Speech
Tianlun Zuo, Ziyu Zhang, Tingzhi Mao +2
cs.CLcs.SDarXiv:2609.11545v12026High Fidelity Video Prediction with Large Stochastic Recurrent Neural Networks
Ruben Villegas, Arkanath Pathak, Harini Kannan +3
cs.CVarXiv:1911.01655v12019Generative AI in the Construction Industry: Opportunities & Challenges
Prashnna Ghimire, Kyungki Kim, Manoj Acharya
cs.AIcs.LGarXiv:2310.04427v12023On Learning Sets of Symmetric Elements
Haggai Maron, Or Litany, Gal Chechik +1
cs.LGstat.MLarXiv:2002.08599v42020Recurrent Attention Models for Depth-Based Person Identification
Albert Haque, Alexandre Alahi, Li Fei-Fei
cs.CVarXiv:1611.07212v12016BERN2: an advanced neural biomedical named entity recognition and normalization tool
Mujeen Sung, Minbyul Jeong, Yonghwa Choi +3
cs.CLarXiv:2201.02080v32022A Training-Free, Alignment-Free Approach to Corporate Intelligence: Application to SEC Filings
Jean-François Delpech
cs.CLarXiv:2609.11620v12026Pre-gated MoE: An Algorithm-System Co-Design for Fast and Scalable Mixture-of-Expert Inference
Ranggi Hwang, Jianyu Wei, Shijie Cao +4
cs.LGcs.AIcs.ARarXiv:2308.12066v32023Optimal Phase Transitions in Compressed Sensing
Yihong Wu, Sergio Verdú
cs.ITmath.STarXiv:1111.6822v22011Weakly-supervised localization of diabetic retinopathy lesions in retinal fundus images
Waleed M. Gondal, Jan M. Köhler, René Grzeszick +2
cs.CVarXiv:1706.09634v12017Structural priors for data-efficient language learning
Yana Veitsman, Jonas Mayer Martins, Jonathan Lautenschlager +1
cs.CLcs.AIcs.LGarXiv:2609.11505v12026A Practical Evaluation of Commercial Industrial Augmented Reality Systems in an Industry 4.0 Shipyard
Oscar Blanco-Novoa, Tiago M Fernandez-Carames, Paula Fraga-Lamas +1
cs.HCarXiv:2402.00925v12024An Analysis of ISO 26262: Using Machine Learning Safely in Automotive Software
Rick Salay, Rodrigo Queiroz, Krzysztof Czarnecki
cs.AIcs.LGcs.SEarXiv:1709.02435v12017DebtRank: A microscopic foundation for shock propagation
Marco Bardoscia, Stefano Battiston, Fabio Caccioli +1
q-fin.RMarXiv:1504.01857v22015SWRouter: Similarity-Contractive Window Routing for Multi-Turn Large Language Model Conversations
Yu Wang, Yuchen Li, Rui Kong +11
cs.CLcs.AIcs.IRarXiv:2609.11414v12026Delay-Based Back-Pressure Scheduling in Multihop Wireless Networks
Bo Ji, Changhee Joo, Ness B. Shroff
cs.NIcs.PFarXiv:1011.5674v32010Context-aware Captions from Context-agnostic Supervision
Ramakrishna Vedantam, Samy Bengio, Kevin Murphy +2
cs.CVcs.AIarXiv:1701.02870v32017Towards On-Device Evidence Gathering for Intimate Partner Infiltration: A Feasibility Study for Joint Identity-Action Detection
Weisi Yang, Shinan Liu, Feng Xiao +2
cs.CRcs.CYcs.HCarXiv:2502.03682v32025FasterViT: Fast Vision Transformers with Hierarchical Attention
Ali Hatamizadeh, Greg Heinrich, Hongxu Yin +4
cs.CVcs.AIcs.LGarXiv:2306.06189v22023Semantics Disentangling for Generalized Zero-Shot Learning
Zhi Chen, Yadan Luo, Ruihong Qiu +4
cs.CVarXiv:2101.07978v52021Underwater Image Enhancement by Transformer-based Diffusion Model with Non-uniform Sampling for Skip Strategy
Yi Tang, Takafumi Iwaguchi, Hiroshi Kawasaki
cs.CVarXiv:2309.03445v12023Active Learning for Deep Object Detection via Probabilistic Modeling
Jiwoong Choi, Ismail Elezi, Hyuk-Jae Lee +2
cs.CVarXiv:2103.16130v22021Creating a Live, Public Short Message Service Corpus: The NUS SMS Corpus
Tao Chen, Min-Yen Kan
cs.CLarXiv:1112.2468v12011Face Recognition: Too Bias, or Not Too Bias?
Joseph P Robinson, Gennady Livitz, Yann Henon +3
cs.CVarXiv:2002.06483v42020TransClean: A Benchmark for Detecting and Extracting Clean Translations from Large Language Model Outputs
Shenbin Qian, Yves Scherrer
cs.CLarXiv:2609.11399v12026E-CONAN (Entailment, CONtradition And Neutral) Benchmarks: Arabic Textual Entailment and Natural Inference Datasets
Khloud AL Jallad, Nada Ghneim, Ghaida Rebdawi
cs.CLcs.AIcs.LGarXiv:2609.11334v12026SEAR: Segment-Evidence-Aware Routing for Weak-to-Strong Multilingual Speech MCQ
Huy Hoang Le, Long-Bao Nguyen, Minh Tri Dao
cs.CLcs.SDarXiv:2609.11355v12026COAST: COntrollable Arbitrary-Sampling NeTwork for Compressive Sensing
Di You, Jian Zhang, Jingfen Xie +2
cs.CVeess.IVarXiv:2107.07225v12021DAG-Recurrent Neural Networks For Scene Labeling
Bing Shuai, Zhen Zuo, Gang Wang +1
cs.CVarXiv:1509.00552v22015Rateless Codes for Near-Perfect Load Balancing in Distributed Matrix-Vector Multiplication
Ankur Mallick, Malhar Chaudhari, Utsav Sheth +2
cs.DCcs.ITarXiv:1804.10331v52018XGBOD: Improving Supervised Outlier Detection with Unsupervised Representation Learning
Yue Zhao, Maciej K. Hryniewicki
cs.LGcs.DBcs.IRarXiv:1912.00290v12019Ultrareliable and Low-Latency Communication Techniques for Tactile Internet Services
Kwang Soon Kim, Dong Ku Kim, Chan-Byoung Chae +9
cs.ITeess.SParXiv:1907.04474v12019Block-Recurrent Transformers
DeLesley Hutchins, Imanol Schlag, Yuhuai Wu +2
cs.LGcs.AIcs.NEarXiv:2203.07852v32022DialogLM: Pre-trained Model for Long Dialogue Understanding and Summarization
Ming Zhong, Yang Liu, Yichong Xu +2
cs.CLarXiv:2109.02492v22021VCT: A Video Compression Transformer
Fabian Mentzer, George Toderici, David Minnen +4
cs.CVcs.LGeess.IVarXiv:2206.07307v22022DeepFilterNet: A Low Complexity Speech Enhancement Framework for Full-Band Audio based on Deep Filtering
Hendrik Schröter, Alberto N. Escalante-B., Tobias Rosenkranz +1
eess.AScs.LGeess.SParXiv:2110.05588v22021Natural SQL: Making SQL Easier to Infer from Natural Language Specifications
Yujian Gan, Xinyun Chen, Jinxia Xie +4
cs.CLarXiv:2109.05153v12021Backdoor Attack in the Physical World
Yiming Li, Tongqing Zhai, Yong Jiang +2
cs.CRcs.AIcs.CVarXiv:2104.02361v22021Performance-Efficiency Trade-offs in Unsupervised Pre-training for Speech Recognition
Felix Wu, Kwangyoun Kim, Jing Pan +3
cs.CLcs.LGcs.SDarXiv:2109.06870v12021