Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,941 to 9,000 of 20,199

  1. RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

    Roberta Raileanu, Tim Rocktäschel

    cs.LGcs.AIarXiv:2002.12292v22020
  2. Learning the Constitutive Behavior of Materials via Neural Operators and Causal Attention: Case Studies in Plasticity and Damage

    Rishabh Arora, Lisa Scheunemann, Tim Brepols +1

    cs.LGcs.CEarXiv:2609.02194v12026
  3. Towards Better Evaluation for Dynamic Link Prediction

    Farimah Poursafaei, Shenyang Huang, Kellin Pelrine +1

    cs.LGcs.SIarXiv:2207.10128v22022
  4. Does Prompt Formatting Have Any Impact on LLM Performance?

    Jia He, Mukund Rungta, David Koleczek +3

    cs.CLcs.LGarXiv:2411.10541v12024
  5. From topology learning to graph generation: A unifying perspective

    Xiaowen Dong, Hoi-To Wai, Siheng Chen +2

    stat.MLcs.LGeess.SParXiv:2609.02286v12026
  6. Molecular geometry prediction using a deep generative graph neural network

    Elman Mansimov, Omar Mahmood, Seokho Kang +1

    cs.LGphysics.comp-phstat.MLarXiv:1904.00314v22019
  7. A Critic Evaluation of Methods for COVID-19 Automatic Detection from X-Ray Images

    Gianluca Maguolo, Loris Nanni

    eess.IVcs.CVcs.LGarXiv:2004.12823v42020
  8. Lingvo: a Modular and Scalable Framework for Sequence-to-Sequence Modeling

    Jonathan Shen, Patrick Nguyen, Yonghui Wu +88

    cs.LGstat.MLarXiv:1902.08295v12019
  9. Towards Continual Knowledge Learning of Language Models

    Joel Jang, Seonghyeon Ye, Sohee Yang +5

    cs.CLcs.LGarXiv:2110.03215v42021
  10. Crystal Structure Prediction by Joint Equivariant Diffusion

    Rui Jiao, Wenbing Huang, Peijia Lin +4

    cond-mat.mtrl-scics.LGarXiv:2309.04475v22023
  11. CBraMod: A Criss-Cross Brain Foundation Model for EEG Decoding

    Jiquan Wang, Sha Zhao, Zhiling Luo +5

    eess.SPcs.AIcs.LGarXiv:2412.07236v62024
  12. An Analysis of the t-SNE Algorithm for Data Visualization

    Sanjeev Arora, Wei Hu, Pravesh K. Kothari

    cs.LGarXiv:1803.01768v22018
  13. Cycle Consistent Adversarial Denoising Network for Multiphase Coronary CT Angiography

    Eunhee Kang, Hyun Jung Koo, Dong Hyun Yang +2

    cs.CVcs.AIcs.LGarXiv:1806.09748v32018
  14. Using Natural Language for Reward Shaping in Reinforcement Learning

    Prasoon Goyal, Scott Niekum, Raymond J. Mooney

    cs.LGcs.AIstat.MLarXiv:1903.02020v22019
  15. Towards a Rigorous Evaluation of XAI Methods on Time Series

    Udo Schlegel, Hiba Arnout, Mennatallah El-Assady +2

    cs.LGcs.AIarXiv:1909.07082v22019
  16. Lower Bounds and Optimal Algorithms for Personalized Federated Learning

    Filip Hanzely, Slavomír Hanzely, Samuel Horváth +1

    cs.LGcs.DCmath.OCarXiv:2010.02372v12020
  17. Generative Adversarial Transformers

    Drew A. Hudson, C. Lawrence Zitnick

    cs.CVcs.AIcs.CLarXiv:2103.01209v42021
  18. Leveraging Recent Advances in Deep Learning for Audio-Visual Emotion Recognition

    Liam Schoneveld, Alice Othmani, Hazem Abdelkawy

    cs.CVcs.LGcs.SDarXiv:2103.09154v22021
  19. C-LSTM: Enabling Efficient LSTM using Structured Compression Techniques on FPGAs

    Shuo Wang, Zhe Li, Caiwen Ding +4

    cs.LGcs.ARarXiv:1803.06305v12018
  20. Alpha-CLIP: A CLIP Model Focusing on Wherever You Want

    Zeyi Sun, Ye Fang, Tong Wu +6

    cs.CVcs.AIcs.CLarXiv:2312.03818v22023
  21. CGMH: Constrained Sentence Generation by Metropolis-Hastings Sampling

    Ning Miao, Hao Zhou, Lili Mou +2

    cs.CLcs.AIcs.LGarXiv:1811.10996v12018
  22. How do World Models and Policies Compose in LLM Agents? A Joint Spectral and Behavioral Account

    Ruize Xu, Xiao Yu, Yujin Tang +2

    cs.LGcs.AIcs.CLarXiv:2608.30067v12026
  23. Deep Reinforcement Learning for Imbalanced Classification

    Enlu Lin, Qiong Chen, Xiaoming Qi

    cs.LGcs.AIstat.MLarXiv:1901.01379v12019
  24. FedSplit: An algorithmic framework for fast federated optimization

    Reese Pathak, Martin J. Wainwright

    cs.LGmath.OCstat.MLarXiv:2005.05238v12020
  25. Deep Online Learning via Meta-Learning: Continual Adaptation for Model-Based RL

    Anusha Nagabandi, Chelsea Finn, Sergey Levine

    cs.LGcs.AIcs.ROarXiv:1812.07671v22018
  26. S-LoRA: Serving Thousands of Concurrent LoRA Adapters

    Ying Sheng, Shiyi Cao, Dacheng Li +9

    cs.LGcs.AIcs.DCarXiv:2311.03285v32023
  27. Janossy Pooling: Learning Deep Permutation-Invariant Functions for Variable-Size Inputs

    Ryan L. Murphy, Balasubramaniam Srinivasan, Vinayak Rao +1

    cs.LGstat.MLarXiv:1811.01900v32018
  28. Harmless interpolation of noisy data in regression

    Vidya Muthukumar, Kailas Vodrahalli, Vignesh Subramanian +1

    cs.LGstat.MLarXiv:1903.09139v22019
  29. Interactive Gibson Benchmark (iGibson 0.5): A Benchmark for Interactive Navigation in Cluttered Environments

    Fei Xia, William B. Shen, Chengshu Li +6

    cs.ROcs.AIcs.CVarXiv:1910.14442v32019
  30. Molecule Edit Graph Attention Network: Modeling Chemical Reactions as Sequences of Graph Edits

    Mikołaj Sacha, Mikołaj Błaż, Piotr Byrski +5

    cs.LGphysics.chem-phstat.MLarXiv:2006.15426v22020
  31. Double Reinforcement Learning for Efficient Off-Policy Evaluation in Markov Decision Processes

    Nathan Kallus, Masatoshi Uehara

    cs.LGcs.AIstat.MLarXiv:1908.08526v32019
  32. Capacity and Trainability in Recurrent Neural Networks

    Jasmine Collins, Jascha Sohl-Dickstein, David Sussillo

    stat.MLcs.AIcs.LGarXiv:1611.09913v32016
  33. Generative and Discriminative Text Classification with Recurrent Neural Networks

    Dani Yogatama, Chris Dyer, Wang Ling +1

    stat.MLcs.CLcs.LGarXiv:1703.01898v22017
  34. Aff-Wild2: Extending the Aff-Wild Database for Affect Recognition

    Dimitrios Kollias, Stefanos Zafeiriou

    cs.CVcs.AIcs.LGarXiv:1811.07770v22018
  35. A General Approach to Adding Differential Privacy to Iterative Training Procedures

    H. Brendan McMahan, Galen Andrew, Ulfar Erlingsson +4

    cs.LGstat.MLarXiv:1812.06210v22018
  36. Perceptually Regularized Diffusion Model for Image Super-Resolution

    Chuxiangbo Wang, Pavithra Venkatachalapathy, Ying Liang +4

    eess.IVcs.CVcs.LGarXiv:2609.02016v12026
  37. EF21: A New, Simpler, Theoretically Better, and Practically Faster Error Feedback

    Peter Richtárik, Igor Sokolov, Ilyas Fatkhullin

    cs.LGmath.OCstat.MLarXiv:2106.05203v12021
  38. Adversarial AutoAugment

    Xinyu Zhang, Qiang Wang, Jian Zhang +1

    cs.CVcs.LGstat.MLarXiv:1912.11188v12019
  39. ICD Coding from Clinical Text Using Multi-Filter Residual Convolutional Neural Network

    Fei Li, Hong Yu

    cs.CLcs.LGarXiv:1912.00862v12019
  40. Train What You Deploy: Closing the MLP Reachability Gap in Low-Rank Clone Distillation

    Wenhui Chen, Zhifeng Li, Jie Zhou +5

    cs.LGcs.CLarXiv:2609.02006v12026
  41. Hierarchical Graph Pooling with Structure Learning

    Zhen Zhang, Jiajun Bu, Martin Ester +4

    cs.LGstat.MLarXiv:1911.05954v32019
  42. DROCC: Deep Robust One-Class Classification

    Sachin Goyal, Aditi Raghunathan, Moksh Jain +2

    cs.LGstat.MLarXiv:2002.12718v22020
  43. On the approximation of functions by tanh neural networks

    Tim De Ryck, Samuel Lanthaler, Siddhartha Mishra

    math.NAcs.LGarXiv:2104.08938v22021
  44. SLCA: Slow Learner with Classifier Alignment for Continual Learning on a Pre-trained Model

    Gengwei Zhang, Liyuan Wang, Guoliang Kang +2

    cs.CVcs.AIcs.LGarXiv:2303.05118v42023
  45. How to Index Item IDs for Recommendation Foundation Models

    Wenyue Hua, Shuyuan Xu, Yingqiang Ge +1

    cs.IRcs.AIcs.CLarXiv:2305.06569v62023
  46. Feature Learning in Infinite-Width Neural Networks

    Greg Yang, Edward J. Hu

    cs.LGcond-mat.dis-nncs.NEarXiv:2011.14522v32020
  47. Do Tabular Foundation Models Know Physics? Contamination, Units, and the Deterministic Limit

    Wassim Tenachi, Yashar Hezaveh, Laurence Perreault Levasseur +1

    cs.LGastro-ph.IMarXiv:2609.02766v12026
  48. TS-CHIEF: A Scalable and Accurate Forest Algorithm for Time Series Classification

    Ahmed Shifaz, Charlotte Pelletier, Francois Petitjean +1

    cs.LGstat.MLarXiv:1906.10329v22019
  49. Convergence for score-based generative modeling with polynomial complexity

    Holden Lee, Jianfeng Lu, Yixin Tan

    cs.LGmath.PRmath.STarXiv:2206.06227v22022
  50. WMLLM: Self-Evolving Optimization Agents via Predict-Then-Act World Modeling

    Zhongzheng Li, Qingsong Ran, Shikun Feng +5

    cs.LGcs.AIarXiv:2609.01608v12026
  51. Graph Prototypical Networks for Few-shot Learning on Attributed Networks

    Kaize Ding, Jianling Wang, Jundong Li +3

    cs.LGcs.SIstat.MLarXiv:2006.12739v32020
  52. Learning Parameterized Skills

    Bruno Da Silva, George Konidaris, Andrew Barto

    cs.LGstat.MLarXiv:1206.6398v22012
  53. Deep Learning of Part-based Representation of Data Using Sparse Autoencoders with Nonnegativity Constraints

    Ehsan Hosseini-Asl, Jacek M. Zurada, Olfa Nasraoui

    cs.LGstat.MLarXiv:1601.02733v12016
  54. QA-LoRA: Quantization-Aware Low-Rank Adaptation of Large Language Models

    Yuhui Xu, Lingxi Xie, Xiaotao Gu +6

    cs.LGcs.CLarXiv:2309.14717v22023
  55. Eliciting ESG Preferences for Reinforcement Learning-Based Portfolio Optimization

    Giovanni Dispoto, Marcello Restelli, Carmine Ventre

    q-fin.PMcs.CEcs.LGarXiv:2609.02677v12026
  56. Harms of Gender Exclusivity and Challenges in Non-Binary Representation in Language Technologies

    Sunipa Dev, Masoud Monajatipoor, Anaelia Ovalle +3

    cs.CLcs.AIcs.LGarXiv:2108.12084v22021
  57. A Realistic Fish-Habitat Dataset to Evaluate Algorithms for Underwater Visual Analysis

    Alzayat Saleh, Issam H. Laradji, Dmitry A. Konovalov +3

    cs.CVcs.LGeess.IVarXiv:2008.12603v12020
  58. FastSecAgg: Scalable Secure Aggregation for Privacy-Preserving Federated Learning

    Swanand Kadhe, Nived Rajaraman, O. Ozan Koyluoglu +1

    cs.CRcs.ITcs.LGarXiv:2009.11248v12020
  59. AutoGCL: Automated Graph Contrastive Learning via Learnable View Generators

    Yihang Yin, Qingzhong Wang, Siyu Huang +2

    cs.LGarXiv:2109.10259v22021
  60. Node-Based Learning of Multiple Gaussian Graphical Models

    Karthik Mohan, Palma London, Maryam Fazel +2

    stat.MLcs.LGmath.OCarXiv:1303.5145v42013