Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,621 to 10,680 of 20,454
Compression-Aware Abstention: Teaching LLMs to Refuse When KV-Compression Masks Remove Answer Evidence
Mohammadali Khodabandehlou, Bhaskar Krishnamachari
cs.CLcs.LGarXiv:2608.29934v12026Re-Identification with Consistent Attentive Siamese Networks
Meng Zheng, Srikrishna Karanam, Ziyan Wu +1
cs.CVcs.LGarXiv:1811.07487v42018MultiRocket: Multiple pooling operators and transformations for fast and effective time series classification
Chang Wei Tan, Angus Dempster, Christoph Bergmeir +1
cs.LGstat.MLarXiv:2102.00457v42021FBCNet: A Multi-view Convolutional Neural Network for Brain-Computer Interface
Ravikiran Mane, Effie Chew, Karen Chua +5
cs.OHcs.AIcs.LGarXiv:2104.01233v12021Augmenting Organizational Decision-Making with Deep Learning Algorithms: Principles, Promises, and Challenges
Yash Raj Shrestha, Vaibhav Krishna, Georg von Krogh
cs.LGarXiv:2011.02834v12020Toward Optimal Feature Selection in Naive Bayes for Text Categorization
Bo Tang, Steven Kay, Haibo He
stat.MLcs.CLcs.IRarXiv:1602.02850v12016Dealing with Non-Stationarity in Multi-Agent Deep Reinforcement Learning
Georgios Papoudakis, Filippos Christianos, Arrasy Rahman +1
cs.LGcs.AIcs.MAarXiv:1906.04737v12019IDNet: Smartphone-based Gait Recognition with Convolutional Neural Networks
Matteo Gadaleta, Michele Rossi
cs.CVcs.LGarXiv:1606.03238v32016Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation
Run Yang, Runpeng Dai, Jie Sun +5
cs.CLcs.LGarXiv:2608.29846v12026Deep Learning for Detecting Building Defects Using Convolutional Neural Networks
Husein Perez, Joseph H. M. Tah, Amir Mosavi
cs.CVcs.AIcs.LGarXiv:1908.04392v12019Taskmaster-1: Toward a Realistic and Diverse Dialog Dataset
Bill Byrne, Karthik Krishnamoorthi, Chinnadhurai Sankar +7
cs.CLcs.AIcs.LGarXiv:1909.05358v12019Description and Discussion on DCASE2020 Challenge Task2: Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
Yuma Koizumi, Yohei Kawaguchi, Keisuke Imoto +8
eess.AScs.LGcs.SDarXiv:2006.05822v22020Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space
Jiechao Xiong, Qing Wang, Zhuoran Yang +7
cs.LGcs.AIstat.MLarXiv:1810.06394v12018Regularization via Mass Transportation
Soroosh Shafieezadeh-Abadeh, Daniel Kuhn, Peyman Mohajerin Esfahani
math.OCcs.LGstat.MLarXiv:1710.10016v32017Language model compression with weighted low-rank factorization
Yen-Chang Hsu, Ting Hua, Sungen Chang +3
cs.LGcs.AIcs.CLarXiv:2207.00112v12022A Systematic Study and Comprehensive Evaluation of ChatGPT on Benchmark Datasets
Md Tahmid Rahman Laskar, M Saiful Bari, Mizanur Rahman +3
cs.CLcs.AIcs.LGarXiv:2305.18486v42023Deep Tracking: Seeing Beyond Seeing Using Recurrent Neural Networks
Peter Ondruska, Ingmar Posner
cs.LGcs.AIcs.CVarXiv:1602.00991v22016Frequency Bias in Neural Networks for Input of Non-Uniform Density
Ronen Basri, Meirav Galun, Amnon Geifman +3
cs.LGstat.MLarXiv:2003.04560v12020MeshDiffusion: Score-based Generative 3D Mesh Modeling
Zhen Liu, Yao Feng, Michael J. Black +3
cs.GRcs.AIcs.CVarXiv:2303.08133v22023Preference Elicitation for Policy Optimization and Application to Aligning Heart Transplantation with Human Values
Itai Zilberstein, Ioannis Anagnostides, Zachary W Sollie +2
cs.AIcs.LGarXiv:2608.28620v12026Provable approximation properties for deep neural networks
Uri Shaham, Alexander Cloninger, Ronald R. Coifman
stat.MLcs.LGcs.NEarXiv:1509.07385v32015Deep Learning-Enabled Semantic Communication Systems with Task-Unaware Transmitter and Dynamic Data
Hongwei Zhang, Shuo Shao, Meixia Tao +2
cs.ITcs.LGcs.NIarXiv:2205.00271v32022Collaborative Filtering in a Non-Uniform World: Learning with the Weighted Trace Norm
Ruslan Salakhutdinov, Nathan Srebro
cs.LGarXiv:1002.2780v12010AirFormer: Predicting Nationwide Air Quality in China with Transformers
Yuxuan Liang, Yutong Xia, Songyu Ke +5
eess.SPcs.LGarXiv:2211.15979v12022Forecast Evaluation for Data Scientists: Common Pitfalls and Best Practices
Hansika Hewamalage, Klaus Ackermann, Christoph Bergmeir
cs.LGstat.MEarXiv:2203.10716v22022Integration of Neural Network-Based Symbolic Regression in Deep Learning for Scientific Discovery
Samuel Kim, Peter Y. Lu, Srijon Mukherjee +4
cs.LGcs.NEphysics.data-anarXiv:1912.04825v22019Deeply AggreVaTeD: Differentiable Imitation Learning for Sequential Prediction
Wen Sun, Arun Venkatraman, Geoffrey J. Gordon +2
cs.LGarXiv:1703.01030v12017Machine Learning-Enhanced Tabu Search for Tactical Wireless Network Design
Wissem Ahmed Zaid, Alain Hertz, Defeng Liu
cs.AIcs.LGmath.COarXiv:2608.28627v12026Structural Pruning for Diffusion Models
Gongfan Fang, Xinyin Ma, Xinchao Wang
cs.LGcs.AIcs.CVarXiv:2305.10924v32023Dynamic Weights in Multi-Objective Deep Reinforcement Learning
Axel Abels, Diederik M. Roijers, Tom Lenaerts +2
cs.LGcs.AIstat.MLarXiv:1809.07803v22018Third-Person Imitation Learning
Bradly C. Stadie, Pieter Abbeel, Ilya Sutskever
cs.LGarXiv:1703.01703v22017Learning Memory Access Patterns
Milad Hashemi, Kevin Swersky, Jamie A. Smith +5
cs.LGstat.MLarXiv:1803.02329v12018AdaPlanner: Adaptive Planning from Feedback with Language Models
Haotian Sun, Yuchen Zhuang, Lingkai Kong +2
cs.CLcs.AIcs.LGarXiv:2305.16653v12023Federated Visual Classification with Real-World Data Distribution
Tzu-Ming Harry Hsu, Hang Qi, Matthew Brown
cs.LGcs.CVstat.MLarXiv:2003.08082v32020Painless Stochastic Gradient: Interpolation, Line-Search, and Convergence Rates
Sharan Vaswani, Aaron Mishkin, Issam Laradji +3
cs.LGmath.OCstat.MLarXiv:1905.09997v52019Learning to Extract Semantic Structure from Documents Using Multimodal Fully Convolutional Neural Network
Xiao Yang, Ersin Yumer, Paul Asente +3
cs.CVcs.LGarXiv:1706.02337v12017Deep autoregressive neural networks for high-dimensional inverse problems in groundwater contaminant source identification
Shaoxing Mo, Nicholas Zabaras, Xiaoqing Shi +1
stat.MLcs.LGarXiv:1812.09444v12018On the Iteration Complexity of Hypergradient Computation
Riccardo Grazzi, Luca Franceschi, Massimiliano Pontil +1
stat.MLcs.LGarXiv:2006.16218v22020Graph Learning based Recommender Systems: A Review
Shoujin Wang, Liang Hu, Yan Wang +6
cs.IRcs.AIcs.LGarXiv:2105.06339v12021A Living Review of Machine Learning for Particle Physics
Matthew Feickert, Benjamin Nachman
hep-phcs.LGhep-exarXiv:2102.02770v12021Selective-Supervised Contrastive Learning with Noisy Labels
Shikun Li, Xiaobo Xia, Shiming Ge +1
cs.CVcs.AIcs.LGarXiv:2203.04181v12022Generating Fact Checking Explanations
Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1
cs.CLcs.AIcs.LGarXiv:2004.05773v12020Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Zhifei Xie, Changqiao Wu
cs.AIcs.CLcs.HCarXiv:2408.16725v32024Beyond Memorization: Violating Privacy Via Inference with Large Language Models
Robin Staab, Mark Vero, Mislav Balunović +1
cs.AIcs.LGarXiv:2310.07298v22023Multi-Modal Hallucination Control by Visual Information Grounding
Alessandro Favero, Luca Zancato, Matthew Trager +5
cs.CVcs.CLcs.LGarXiv:2403.14003v12024Randomized Smoothing of All Shapes and Sizes
Greg Yang, Tony Duan, J. Edward Hu +3
cs.LGcs.CVcs.NEarXiv:2002.08118v52020Theoretical Foundations of t-SNE for Visualizing High-Dimensional Clustered Data
T. Tony Cai, Rong Ma
stat.MLcs.LGmath.STarXiv:2105.07536v42021A Review of Large Language Models and Autonomous Agents in Chemistry
Mayk Caldas Ramos, Christopher J. Collison, Andrew D. White
cs.LGcs.AIcs.CLarXiv:2407.01603v32024Cosine Normalization: Using Cosine Similarity Instead of Dot Product in Neural Networks
Chunjie Luo, Jianfeng Zhan, Lei Wang +1
cs.LGcs.AIstat.MLarXiv:1702.05870v52017Learning to Utilize Shaping Rewards: A New Approach of Reward Shaping
Yujing Hu, Weixun Wang, Hangtian Jia +5
cs.LGcs.AIarXiv:2011.02669v12020Deep learning versus kernel learning: an empirical study of loss landscape geometry and the time evolution of the Neural Tangent Kernel
Stanislav Fort, Gintare Karolina Dziugaite, Mansheej Paul +3
cs.LGstat.MLarXiv:2010.15110v12020Federated Learning for Computational Pathology on Gigapixel Whole Slide Images
Ming Y. Lu, Dehan Kong, Jana Lipkova +5
eess.IVcs.CVcs.LGarXiv:2009.10190v22020Understanding Membership Inferences on Well-Generalized Learning Models
Yunhui Long, Vincent Bindschaedler, Lei Wang +5
cs.CRcs.LGstat.MLarXiv:1802.04889v12018Rearrangement: A Challenge for Embodied AI
Dhruv Batra, Angel X. Chang, Sonia Chernova +9
cs.AIcs.CVcs.LGarXiv:2011.01975v12020Gmail Smart Compose: Real-Time Assisted Writing
Mia Xu Chen, Benjamin N Lee, Gagan Bansal +9
cs.CLcs.LGarXiv:1906.00080v12019MedMamba: Vision Mamba for Medical Image Classification
Yubiao Yue, Zhenzhang Li
eess.IVcs.CVcs.LGarXiv:2403.03849v52024Nested Hierarchical Dirichlet Processes
John Paisley, Chong Wang, David M. Blei +1
stat.MLcs.LGarXiv:1210.6738v42012Tensor Canonical Correlation Analysis for Multi-view Dimension Reduction
Yong Luo, Dacheng Tao, Yonggang Wen +2
stat.MLcs.CVcs.LGarXiv:1502.02330v12015Deep Multimodal Learning for Audio-Visual Speech Recognition
Youssef Mroueh, Etienne Marcheret, Vaibhava Goel
cs.CLcs.LGarXiv:1501.05396v12015A note on the triangle inequality for the Jaccard distance
Sven Kosub
cs.DMcs.IRcs.LGarXiv:1612.02696v12016