Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,621 to 10,680 of 20,454

  1. Compression-Aware Abstention: Teaching LLMs to Refuse When KV-Compression Masks Remove Answer Evidence

    Mohammadali Khodabandehlou, Bhaskar Krishnamachari

    cs.CLcs.LGarXiv:2608.29934v12026
  2. Re-Identification with Consistent Attentive Siamese Networks

    Meng Zheng, Srikrishna Karanam, Ziyan Wu +1

    cs.CVcs.LGarXiv:1811.07487v42018
  3. MultiRocket: Multiple pooling operators and transformations for fast and effective time series classification

    Chang Wei Tan, Angus Dempster, Christoph Bergmeir +1

    cs.LGstat.MLarXiv:2102.00457v42021
  4. FBCNet: A Multi-view Convolutional Neural Network for Brain-Computer Interface

    Ravikiran Mane, Effie Chew, Karen Chua +5

    cs.OHcs.AIcs.LGarXiv:2104.01233v12021
  5. Augmenting Organizational Decision-Making with Deep Learning Algorithms: Principles, Promises, and Challenges

    Yash Raj Shrestha, Vaibhav Krishna, Georg von Krogh

    cs.LGarXiv:2011.02834v12020
  6. Toward Optimal Feature Selection in Naive Bayes for Text Categorization

    Bo Tang, Steven Kay, Haibo He

    stat.MLcs.CLcs.IRarXiv:1602.02850v12016
  7. Dealing with Non-Stationarity in Multi-Agent Deep Reinforcement Learning

    Georgios Papoudakis, Filippos Christianos, Arrasy Rahman +1

    cs.LGcs.AIcs.MAarXiv:1906.04737v12019
  8. IDNet: Smartphone-based Gait Recognition with Convolutional Neural Networks

    Matteo Gadaleta, Michele Rossi

    cs.CVcs.LGarXiv:1606.03238v32016
  9. Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation

    Run Yang, Runpeng Dai, Jie Sun +5

    cs.CLcs.LGarXiv:2608.29846v12026
  10. Deep Learning for Detecting Building Defects Using Convolutional Neural Networks

    Husein Perez, Joseph H. M. Tah, Amir Mosavi

    cs.CVcs.AIcs.LGarXiv:1908.04392v12019
  11. Taskmaster-1: Toward a Realistic and Diverse Dialog Dataset

    Bill Byrne, Karthik Krishnamoorthi, Chinnadhurai Sankar +7

    cs.CLcs.AIcs.LGarXiv:1909.05358v12019
  12. Description and Discussion on DCASE2020 Challenge Task2: Unsupervised Anomalous Sound Detection for Machine Condition Monitoring

    Yuma Koizumi, Yohei Kawaguchi, Keisuke Imoto +8

    eess.AScs.LGcs.SDarXiv:2006.05822v22020
  13. Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space

    Jiechao Xiong, Qing Wang, Zhuoran Yang +7

    cs.LGcs.AIstat.MLarXiv:1810.06394v12018
  14. Regularization via Mass Transportation

    Soroosh Shafieezadeh-Abadeh, Daniel Kuhn, Peyman Mohajerin Esfahani

    math.OCcs.LGstat.MLarXiv:1710.10016v32017
  15. Language model compression with weighted low-rank factorization

    Yen-Chang Hsu, Ting Hua, Sungen Chang +3

    cs.LGcs.AIcs.CLarXiv:2207.00112v12022
  16. A Systematic Study and Comprehensive Evaluation of ChatGPT on Benchmark Datasets

    Md Tahmid Rahman Laskar, M Saiful Bari, Mizanur Rahman +3

    cs.CLcs.AIcs.LGarXiv:2305.18486v42023
  17. Deep Tracking: Seeing Beyond Seeing Using Recurrent Neural Networks

    Peter Ondruska, Ingmar Posner

    cs.LGcs.AIcs.CVarXiv:1602.00991v22016
  18. Frequency Bias in Neural Networks for Input of Non-Uniform Density

    Ronen Basri, Meirav Galun, Amnon Geifman +3

    cs.LGstat.MLarXiv:2003.04560v12020
  19. MeshDiffusion: Score-based Generative 3D Mesh Modeling

    Zhen Liu, Yao Feng, Michael J. Black +3

    cs.GRcs.AIcs.CVarXiv:2303.08133v22023
  20. Preference Elicitation for Policy Optimization and Application to Aligning Heart Transplantation with Human Values

    Itai Zilberstein, Ioannis Anagnostides, Zachary W Sollie +2

    cs.AIcs.LGarXiv:2608.28620v12026
  21. Provable approximation properties for deep neural networks

    Uri Shaham, Alexander Cloninger, Ronald R. Coifman

    stat.MLcs.LGcs.NEarXiv:1509.07385v32015
  22. Deep Learning-Enabled Semantic Communication Systems with Task-Unaware Transmitter and Dynamic Data

    Hongwei Zhang, Shuo Shao, Meixia Tao +2

    cs.ITcs.LGcs.NIarXiv:2205.00271v32022
  23. Collaborative Filtering in a Non-Uniform World: Learning with the Weighted Trace Norm

    Ruslan Salakhutdinov, Nathan Srebro

    cs.LGarXiv:1002.2780v12010
  24. AirFormer: Predicting Nationwide Air Quality in China with Transformers

    Yuxuan Liang, Yutong Xia, Songyu Ke +5

    eess.SPcs.LGarXiv:2211.15979v12022
  25. Forecast Evaluation for Data Scientists: Common Pitfalls and Best Practices

    Hansika Hewamalage, Klaus Ackermann, Christoph Bergmeir

    cs.LGstat.MEarXiv:2203.10716v22022
  26. Integration of Neural Network-Based Symbolic Regression in Deep Learning for Scientific Discovery

    Samuel Kim, Peter Y. Lu, Srijon Mukherjee +4

    cs.LGcs.NEphysics.data-anarXiv:1912.04825v22019
  27. Deeply AggreVaTeD: Differentiable Imitation Learning for Sequential Prediction

    Wen Sun, Arun Venkatraman, Geoffrey J. Gordon +2

    cs.LGarXiv:1703.01030v12017
  28. Machine Learning-Enhanced Tabu Search for Tactical Wireless Network Design

    Wissem Ahmed Zaid, Alain Hertz, Defeng Liu

    cs.AIcs.LGmath.COarXiv:2608.28627v12026
  29. Structural Pruning for Diffusion Models

    Gongfan Fang, Xinyin Ma, Xinchao Wang

    cs.LGcs.AIcs.CVarXiv:2305.10924v32023
  30. Dynamic Weights in Multi-Objective Deep Reinforcement Learning

    Axel Abels, Diederik M. Roijers, Tom Lenaerts +2

    cs.LGcs.AIstat.MLarXiv:1809.07803v22018
  31. Third-Person Imitation Learning

    Bradly C. Stadie, Pieter Abbeel, Ilya Sutskever

    cs.LGarXiv:1703.01703v22017
  32. Learning Memory Access Patterns

    Milad Hashemi, Kevin Swersky, Jamie A. Smith +5

    cs.LGstat.MLarXiv:1803.02329v12018
  33. AdaPlanner: Adaptive Planning from Feedback with Language Models

    Haotian Sun, Yuchen Zhuang, Lingkai Kong +2

    cs.CLcs.AIcs.LGarXiv:2305.16653v12023
  34. Federated Visual Classification with Real-World Data Distribution

    Tzu-Ming Harry Hsu, Hang Qi, Matthew Brown

    cs.LGcs.CVstat.MLarXiv:2003.08082v32020
  35. Painless Stochastic Gradient: Interpolation, Line-Search, and Convergence Rates

    Sharan Vaswani, Aaron Mishkin, Issam Laradji +3

    cs.LGmath.OCstat.MLarXiv:1905.09997v52019
  36. Learning to Extract Semantic Structure from Documents Using Multimodal Fully Convolutional Neural Network

    Xiao Yang, Ersin Yumer, Paul Asente +3

    cs.CVcs.LGarXiv:1706.02337v12017
  37. Deep autoregressive neural networks for high-dimensional inverse problems in groundwater contaminant source identification

    Shaoxing Mo, Nicholas Zabaras, Xiaoqing Shi +1

    stat.MLcs.LGarXiv:1812.09444v12018
  38. On the Iteration Complexity of Hypergradient Computation

    Riccardo Grazzi, Luca Franceschi, Massimiliano Pontil +1

    stat.MLcs.LGarXiv:2006.16218v22020
  39. Graph Learning based Recommender Systems: A Review

    Shoujin Wang, Liang Hu, Yan Wang +6

    cs.IRcs.AIcs.LGarXiv:2105.06339v12021
  40. A Living Review of Machine Learning for Particle Physics

    Matthew Feickert, Benjamin Nachman

    hep-phcs.LGhep-exarXiv:2102.02770v12021
  41. Selective-Supervised Contrastive Learning with Noisy Labels

    Shikun Li, Xiaobo Xia, Shiming Ge +1

    cs.CVcs.AIcs.LGarXiv:2203.04181v12022
  42. Generating Fact Checking Explanations

    Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1

    cs.CLcs.AIcs.LGarXiv:2004.05773v12020
  43. Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming

    Zhifei Xie, Changqiao Wu

    cs.AIcs.CLcs.HCarXiv:2408.16725v32024
  44. Beyond Memorization: Violating Privacy Via Inference with Large Language Models

    Robin Staab, Mark Vero, Mislav Balunović +1

    cs.AIcs.LGarXiv:2310.07298v22023
  45. Multi-Modal Hallucination Control by Visual Information Grounding

    Alessandro Favero, Luca Zancato, Matthew Trager +5

    cs.CVcs.CLcs.LGarXiv:2403.14003v12024
  46. Randomized Smoothing of All Shapes and Sizes

    Greg Yang, Tony Duan, J. Edward Hu +3

    cs.LGcs.CVcs.NEarXiv:2002.08118v52020
  47. Theoretical Foundations of t-SNE for Visualizing High-Dimensional Clustered Data

    T. Tony Cai, Rong Ma

    stat.MLcs.LGmath.STarXiv:2105.07536v42021
  48. A Review of Large Language Models and Autonomous Agents in Chemistry

    Mayk Caldas Ramos, Christopher J. Collison, Andrew D. White

    cs.LGcs.AIcs.CLarXiv:2407.01603v32024
  49. Cosine Normalization: Using Cosine Similarity Instead of Dot Product in Neural Networks

    Chunjie Luo, Jianfeng Zhan, Lei Wang +1

    cs.LGcs.AIstat.MLarXiv:1702.05870v52017
  50. Learning to Utilize Shaping Rewards: A New Approach of Reward Shaping

    Yujing Hu, Weixun Wang, Hangtian Jia +5

    cs.LGcs.AIarXiv:2011.02669v12020
  51. Deep learning versus kernel learning: an empirical study of loss landscape geometry and the time evolution of the Neural Tangent Kernel

    Stanislav Fort, Gintare Karolina Dziugaite, Mansheej Paul +3

    cs.LGstat.MLarXiv:2010.15110v12020
  52. Federated Learning for Computational Pathology on Gigapixel Whole Slide Images

    Ming Y. Lu, Dehan Kong, Jana Lipkova +5

    eess.IVcs.CVcs.LGarXiv:2009.10190v22020
  53. Understanding Membership Inferences on Well-Generalized Learning Models

    Yunhui Long, Vincent Bindschaedler, Lei Wang +5

    cs.CRcs.LGstat.MLarXiv:1802.04889v12018
  54. Rearrangement: A Challenge for Embodied AI

    Dhruv Batra, Angel X. Chang, Sonia Chernova +9

    cs.AIcs.CVcs.LGarXiv:2011.01975v12020
  55. Gmail Smart Compose: Real-Time Assisted Writing

    Mia Xu Chen, Benjamin N Lee, Gagan Bansal +9

    cs.CLcs.LGarXiv:1906.00080v12019
  56. MedMamba: Vision Mamba for Medical Image Classification

    Yubiao Yue, Zhenzhang Li

    eess.IVcs.CVcs.LGarXiv:2403.03849v52024
  57. Nested Hierarchical Dirichlet Processes

    John Paisley, Chong Wang, David M. Blei +1

    stat.MLcs.LGarXiv:1210.6738v42012
  58. Tensor Canonical Correlation Analysis for Multi-view Dimension Reduction

    Yong Luo, Dacheng Tao, Yonggang Wen +2

    stat.MLcs.CVcs.LGarXiv:1502.02330v12015
  59. Deep Multimodal Learning for Audio-Visual Speech Recognition

    Youssef Mroueh, Etienne Marcheret, Vaibhava Goel

    cs.CLcs.LGarXiv:1501.05396v12015
  60. A note on the triangle inequality for the Jaccard distance

    Sven Kosub

    cs.DMcs.IRcs.LGarXiv:1612.02696v12016