Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

10,381 to 10,440 of 20,193

  1. Dynamic Weights in Multi-Objective Deep Reinforcement Learning

    Axel Abels, Diederik M. Roijers, Tom Lenaerts +2

    cs.LGcs.AIstat.MLarXiv:1809.07803v22018
  2. Third-Person Imitation Learning

    Bradly C. Stadie, Pieter Abbeel, Ilya Sutskever

    cs.LGarXiv:1703.01703v22017
  3. Learning Memory Access Patterns

    Milad Hashemi, Kevin Swersky, Jamie A. Smith +5

    cs.LGstat.MLarXiv:1803.02329v12018
  4. AdaPlanner: Adaptive Planning from Feedback with Language Models

    Haotian Sun, Yuchen Zhuang, Lingkai Kong +2

    cs.CLcs.AIcs.LGarXiv:2305.16653v12023
  5. Federated Visual Classification with Real-World Data Distribution

    Tzu-Ming Harry Hsu, Hang Qi, Matthew Brown

    cs.LGcs.CVstat.MLarXiv:2003.08082v32020
  6. Painless Stochastic Gradient: Interpolation, Line-Search, and Convergence Rates

    Sharan Vaswani, Aaron Mishkin, Issam Laradji +3

    cs.LGmath.OCstat.MLarXiv:1905.09997v52019
  7. Learning to Extract Semantic Structure from Documents Using Multimodal Fully Convolutional Neural Network

    Xiao Yang, Ersin Yumer, Paul Asente +3

    cs.CVcs.LGarXiv:1706.02337v12017
  8. Deep autoregressive neural networks for high-dimensional inverse problems in groundwater contaminant source identification

    Shaoxing Mo, Nicholas Zabaras, Xiaoqing Shi +1

    stat.MLcs.LGarXiv:1812.09444v12018
  9. On the Iteration Complexity of Hypergradient Computation

    Riccardo Grazzi, Luca Franceschi, Massimiliano Pontil +1

    stat.MLcs.LGarXiv:2006.16218v22020
  10. Graph Learning based Recommender Systems: A Review

    Shoujin Wang, Liang Hu, Yan Wang +6

    cs.IRcs.AIcs.LGarXiv:2105.06339v12021
  11. A Living Review of Machine Learning for Particle Physics

    Matthew Feickert, Benjamin Nachman

    hep-phcs.LGhep-exarXiv:2102.02770v12021
  12. Selective-Supervised Contrastive Learning with Noisy Labels

    Shikun Li, Xiaobo Xia, Shiming Ge +1

    cs.CVcs.AIcs.LGarXiv:2203.04181v12022
  13. Generating Fact Checking Explanations

    Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1

    cs.CLcs.AIcs.LGarXiv:2004.05773v12020
  14. Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming

    Zhifei Xie, Changqiao Wu

    cs.AIcs.CLcs.HCarXiv:2408.16725v32024
  15. Beyond Memorization: Violating Privacy Via Inference with Large Language Models

    Robin Staab, Mark Vero, Mislav Balunović +1

    cs.AIcs.LGarXiv:2310.07298v22023
  16. Multi-Modal Hallucination Control by Visual Information Grounding

    Alessandro Favero, Luca Zancato, Matthew Trager +5

    cs.CVcs.CLcs.LGarXiv:2403.14003v12024
  17. Randomized Smoothing of All Shapes and Sizes

    Greg Yang, Tony Duan, J. Edward Hu +3

    cs.LGcs.CVcs.NEarXiv:2002.08118v52020
  18. Theoretical Foundations of t-SNE for Visualizing High-Dimensional Clustered Data

    T. Tony Cai, Rong Ma

    stat.MLcs.LGmath.STarXiv:2105.07536v42021
  19. A Review of Large Language Models and Autonomous Agents in Chemistry

    Mayk Caldas Ramos, Christopher J. Collison, Andrew D. White

    cs.LGcs.AIcs.CLarXiv:2407.01603v32024
  20. Cosine Normalization: Using Cosine Similarity Instead of Dot Product in Neural Networks

    Chunjie Luo, Jianfeng Zhan, Lei Wang +1

    cs.LGcs.AIstat.MLarXiv:1702.05870v52017
  21. Learning to Utilize Shaping Rewards: A New Approach of Reward Shaping

    Yujing Hu, Weixun Wang, Hangtian Jia +5

    cs.LGcs.AIarXiv:2011.02669v12020
  22. Deep learning versus kernel learning: an empirical study of loss landscape geometry and the time evolution of the Neural Tangent Kernel

    Stanislav Fort, Gintare Karolina Dziugaite, Mansheej Paul +3

    cs.LGstat.MLarXiv:2010.15110v12020
  23. Federated Learning for Computational Pathology on Gigapixel Whole Slide Images

    Ming Y. Lu, Dehan Kong, Jana Lipkova +5

    eess.IVcs.CVcs.LGarXiv:2009.10190v22020
  24. Understanding Membership Inferences on Well-Generalized Learning Models

    Yunhui Long, Vincent Bindschaedler, Lei Wang +5

    cs.CRcs.LGstat.MLarXiv:1802.04889v12018
  25. Rearrangement: A Challenge for Embodied AI

    Dhruv Batra, Angel X. Chang, Sonia Chernova +9

    cs.AIcs.CVcs.LGarXiv:2011.01975v12020
  26. Gmail Smart Compose: Real-Time Assisted Writing

    Mia Xu Chen, Benjamin N Lee, Gagan Bansal +9

    cs.CLcs.LGarXiv:1906.00080v12019
  27. MedMamba: Vision Mamba for Medical Image Classification

    Yubiao Yue, Zhenzhang Li

    eess.IVcs.CVcs.LGarXiv:2403.03849v52024
  28. Nested Hierarchical Dirichlet Processes

    John Paisley, Chong Wang, David M. Blei +1

    stat.MLcs.LGarXiv:1210.6738v42012
  29. Tensor Canonical Correlation Analysis for Multi-view Dimension Reduction

    Yong Luo, Dacheng Tao, Yonggang Wen +2

    stat.MLcs.CVcs.LGarXiv:1502.02330v12015
  30. Deep Multimodal Learning for Audio-Visual Speech Recognition

    Youssef Mroueh, Etienne Marcheret, Vaibhava Goel

    cs.CLcs.LGarXiv:1501.05396v12015
  31. A note on the triangle inequality for the Jaccard distance

    Sven Kosub

    cs.DMcs.IRcs.LGarXiv:1612.02696v12016
  32. Machine Learning-Based Prototyping of Graphical User Interfaces for Mobile Apps

    Kevin Moran, Carlos Bernal-Cárdenas, Michael Curcio +2

    cs.SEcs.CVcs.LGarXiv:1802.02312v22018
  33. Multi-Scale High-Resolution Vision Transformer for Semantic Segmentation

    Jiaqi Gu, Hyoukjun Kwon, Dilin Wang +6

    cs.CVcs.AIcs.LGarXiv:2111.01236v22021
  34. Multi-View Spatial-Temporal Graph Convolutional Networks with Domain Generalization for Sleep Stage Classification

    Ziyu Jia, Youfang Lin, Jing Wang +5

    eess.SPcs.AIcs.CVarXiv:2109.01824v12021
  35. DIVA: Domain Invariant Variational Autoencoders

    Maximilian Ilse, Jakub M. Tomczak, Christos Louizos +1

    stat.MLcs.LGarXiv:1905.10427v22019
  36. On the Origin of Implicit Regularization in Stochastic Gradient Descent

    Samuel L. Smith, Benoit Dherin, David G. T. Barrett +1

    cs.LGstat.MLarXiv:2101.12176v12021
  37. Adversarial Filters of Dataset Biases

    Ronan Le Bras, Swabha Swayamdipta, Chandra Bhagavatula +4

    cs.LGcs.AIcs.CLarXiv:2002.04108v32020
  38. Deep Gaussian Processes for Regression using Approximate Expectation Propagation

    Thang D. Bui, Daniel Hernández-Lobato, Yingzhen Li +2

    stat.MLcs.LGarXiv:1602.04133v12016
  39. Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees

    Yuping Luo, Huazhe Xu, Yuanzhi Li +3

    cs.LGcs.AIstat.MLarXiv:1807.03858v52018
  40. 2 OLMo 2 Furious

    Team OLMo, Pete Walsh, Luca Soldaini +40

    cs.CLcs.LGarXiv:2501.00656v32024
  41. The Unreasonable Ineffectiveness of the Deeper Layers

    Andrey Gromov, Kushal Tirumala, Hassan Shapourian +2

    cs.CLcs.LGstat.MLarXiv:2403.17887v22024
  42. Learning to cluster in order to transfer across domains and tasks

    Yen-Chang Hsu, Zhaoyang Lv, Zsolt Kira

    cs.LGcs.AIcs.CVarXiv:1711.10125v32017
  43. Progressive Prompts: Continual Learning for Language Models

    Anastasia Razdaibiedina, Yuning Mao, Rui Hou +3

    cs.CLcs.AIcs.LGarXiv:2301.12314v12023
  44. Defending Against Indirect Prompt Injection Attacks With Spotlighting

    Keegan Hines, Gary Lopez, Matthew Hall +3

    cs.CRcs.CLcs.LGarXiv:2403.14720v12024
  45. Autoregressive Diffusion Models

    Emiel Hoogeboom, Alexey A. Gritsenko, Jasmijn Bastings +3

    cs.LGstat.MLarXiv:2110.02037v22021
  46. FLamby: Datasets and Benchmarks for Cross-Silo Federated Learning in Realistic Healthcare Settings

    Jean Ogier du Terrail, Samy-Safwan Ayed, Edwige Cyffers +21

    cs.LGcs.CVarXiv:2210.04620v32022
  47. Combined Scaling for Zero-shot Transfer Learning

    Hieu Pham, Zihang Dai, Golnaz Ghiasi +9

    cs.LGcs.CLcs.CVarXiv:2111.10050v32021
  48. ViNT: A Foundation Model for Visual Navigation

    Dhruv Shah, Ajay Sridhar, Nitish Dashora +4

    cs.ROcs.CVcs.LGarXiv:2306.14846v22023
  49. Variable Impedance Control in End-Effector Space: An Action Space for Reinforcement Learning in Contact-Rich Tasks

    Roberto Martín-Martín, Michelle A. Lee, Rachel Gardner +3

    cs.ROcs.AIcs.LGarXiv:1906.08880v22019
  50. TACS: Trajectory-Aware Candidate Selection for LLM Jailbreak Suffix Optimization

    Shiliang Xiao

    cs.CLcs.LGarXiv:2608.29564v12026
  51. Field-weighted Factorization Machines for Click-Through Rate Prediction in Display Advertising

    Junwei Pan, Jian Xu, Alfonso Lobos Ruiz +4

    cs.LGstat.MLarXiv:1806.03514v22018
  52. Detection of Novel Social Bots by Ensembles of Specialized Classifiers

    Mohsen Sayyadiharikandeh, Onur Varol, Kai-Cheng Yang +2

    cs.SIcs.IRcs.LGarXiv:2006.06867v22020
  53. CodeNeRF: Disentangled Neural Radiance Fields for Object Categories

    Wonbong Jang, Lourdes Agapito

    cs.GRcs.CVcs.LGarXiv:2109.01750v12021
  54. Structured Pruning Learns Compact and Accurate Models

    Mengzhou Xia, Zexuan Zhong, Danqi Chen

    cs.CLcs.LGarXiv:2204.00408v32022
  55. LoGo: Token-Level Dynamic Local-Global Attention

    Yuqi Pan, Zheng Li, Bohao Tang +2

    cs.CLcs.LGarXiv:2608.29539v12026
  56. DarkneTZ: Towards Model Privacy at the Edge using Trusted Execution Environments

    Fan Mo, Ali Shahin Shamsabadi, Kleomenis Katevas +4

    cs.LGcs.CRstat.MLarXiv:2004.05703v12020
  57. DeepSight: Mitigating Backdoor Attacks in Federated Learning Through Deep Model Inspection

    Phillip Rieger, Thien Duc Nguyen, Markus Miettinen +1

    cs.CRcs.LGarXiv:2201.00763v12022
  58. Predicting Head Movement in Panoramic Video: A Deep Reinforcement Learning Approach

    Yuhang Song, Mai Xu, Jianyi Wang +3

    cs.CVcs.LGarXiv:1710.10755v52017
  59. QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving

    Yujun Lin, Haotian Tang, Shang Yang +4

    cs.CLcs.AIcs.LGarXiv:2405.04532v32024
  60. A Simple Convergence Proof of Adam and Adagrad

    Alexandre Défossez, Léon Bottou, Francis Bach +1

    stat.MLcs.LGarXiv:2003.02395v32020