Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
10,381 to 10,440 of 20,193
Dynamic Weights in Multi-Objective Deep Reinforcement Learning
Axel Abels, Diederik M. Roijers, Tom Lenaerts +2
cs.LGcs.AIstat.MLarXiv:1809.07803v22018Third-Person Imitation Learning
Bradly C. Stadie, Pieter Abbeel, Ilya Sutskever
cs.LGarXiv:1703.01703v22017Learning Memory Access Patterns
Milad Hashemi, Kevin Swersky, Jamie A. Smith +5
cs.LGstat.MLarXiv:1803.02329v12018AdaPlanner: Adaptive Planning from Feedback with Language Models
Haotian Sun, Yuchen Zhuang, Lingkai Kong +2
cs.CLcs.AIcs.LGarXiv:2305.16653v12023Federated Visual Classification with Real-World Data Distribution
Tzu-Ming Harry Hsu, Hang Qi, Matthew Brown
cs.LGcs.CVstat.MLarXiv:2003.08082v32020Painless Stochastic Gradient: Interpolation, Line-Search, and Convergence Rates
Sharan Vaswani, Aaron Mishkin, Issam Laradji +3
cs.LGmath.OCstat.MLarXiv:1905.09997v52019Learning to Extract Semantic Structure from Documents Using Multimodal Fully Convolutional Neural Network
Xiao Yang, Ersin Yumer, Paul Asente +3
cs.CVcs.LGarXiv:1706.02337v12017Deep autoregressive neural networks for high-dimensional inverse problems in groundwater contaminant source identification
Shaoxing Mo, Nicholas Zabaras, Xiaoqing Shi +1
stat.MLcs.LGarXiv:1812.09444v12018On the Iteration Complexity of Hypergradient Computation
Riccardo Grazzi, Luca Franceschi, Massimiliano Pontil +1
stat.MLcs.LGarXiv:2006.16218v22020Graph Learning based Recommender Systems: A Review
Shoujin Wang, Liang Hu, Yan Wang +6
cs.IRcs.AIcs.LGarXiv:2105.06339v12021A Living Review of Machine Learning for Particle Physics
Matthew Feickert, Benjamin Nachman
hep-phcs.LGhep-exarXiv:2102.02770v12021Selective-Supervised Contrastive Learning with Noisy Labels
Shikun Li, Xiaobo Xia, Shiming Ge +1
cs.CVcs.AIcs.LGarXiv:2203.04181v12022Generating Fact Checking Explanations
Pepa Atanasova, Jakob Grue Simonsen, Christina Lioma +1
cs.CLcs.AIcs.LGarXiv:2004.05773v12020Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Zhifei Xie, Changqiao Wu
cs.AIcs.CLcs.HCarXiv:2408.16725v32024Beyond Memorization: Violating Privacy Via Inference with Large Language Models
Robin Staab, Mark Vero, Mislav Balunović +1
cs.AIcs.LGarXiv:2310.07298v22023Multi-Modal Hallucination Control by Visual Information Grounding
Alessandro Favero, Luca Zancato, Matthew Trager +5
cs.CVcs.CLcs.LGarXiv:2403.14003v12024Randomized Smoothing of All Shapes and Sizes
Greg Yang, Tony Duan, J. Edward Hu +3
cs.LGcs.CVcs.NEarXiv:2002.08118v52020Theoretical Foundations of t-SNE for Visualizing High-Dimensional Clustered Data
T. Tony Cai, Rong Ma
stat.MLcs.LGmath.STarXiv:2105.07536v42021A Review of Large Language Models and Autonomous Agents in Chemistry
Mayk Caldas Ramos, Christopher J. Collison, Andrew D. White
cs.LGcs.AIcs.CLarXiv:2407.01603v32024Cosine Normalization: Using Cosine Similarity Instead of Dot Product in Neural Networks
Chunjie Luo, Jianfeng Zhan, Lei Wang +1
cs.LGcs.AIstat.MLarXiv:1702.05870v52017Learning to Utilize Shaping Rewards: A New Approach of Reward Shaping
Yujing Hu, Weixun Wang, Hangtian Jia +5
cs.LGcs.AIarXiv:2011.02669v12020Deep learning versus kernel learning: an empirical study of loss landscape geometry and the time evolution of the Neural Tangent Kernel
Stanislav Fort, Gintare Karolina Dziugaite, Mansheej Paul +3
cs.LGstat.MLarXiv:2010.15110v12020Federated Learning for Computational Pathology on Gigapixel Whole Slide Images
Ming Y. Lu, Dehan Kong, Jana Lipkova +5
eess.IVcs.CVcs.LGarXiv:2009.10190v22020Understanding Membership Inferences on Well-Generalized Learning Models
Yunhui Long, Vincent Bindschaedler, Lei Wang +5
cs.CRcs.LGstat.MLarXiv:1802.04889v12018Rearrangement: A Challenge for Embodied AI
Dhruv Batra, Angel X. Chang, Sonia Chernova +9
cs.AIcs.CVcs.LGarXiv:2011.01975v12020Gmail Smart Compose: Real-Time Assisted Writing
Mia Xu Chen, Benjamin N Lee, Gagan Bansal +9
cs.CLcs.LGarXiv:1906.00080v12019MedMamba: Vision Mamba for Medical Image Classification
Yubiao Yue, Zhenzhang Li
eess.IVcs.CVcs.LGarXiv:2403.03849v52024Nested Hierarchical Dirichlet Processes
John Paisley, Chong Wang, David M. Blei +1
stat.MLcs.LGarXiv:1210.6738v42012Tensor Canonical Correlation Analysis for Multi-view Dimension Reduction
Yong Luo, Dacheng Tao, Yonggang Wen +2
stat.MLcs.CVcs.LGarXiv:1502.02330v12015Deep Multimodal Learning for Audio-Visual Speech Recognition
Youssef Mroueh, Etienne Marcheret, Vaibhava Goel
cs.CLcs.LGarXiv:1501.05396v12015A note on the triangle inequality for the Jaccard distance
Sven Kosub
cs.DMcs.IRcs.LGarXiv:1612.02696v12016Machine Learning-Based Prototyping of Graphical User Interfaces for Mobile Apps
Kevin Moran, Carlos Bernal-Cárdenas, Michael Curcio +2
cs.SEcs.CVcs.LGarXiv:1802.02312v22018Multi-Scale High-Resolution Vision Transformer for Semantic Segmentation
Jiaqi Gu, Hyoukjun Kwon, Dilin Wang +6
cs.CVcs.AIcs.LGarXiv:2111.01236v22021Multi-View Spatial-Temporal Graph Convolutional Networks with Domain Generalization for Sleep Stage Classification
Ziyu Jia, Youfang Lin, Jing Wang +5
eess.SPcs.AIcs.CVarXiv:2109.01824v12021DIVA: Domain Invariant Variational Autoencoders
Maximilian Ilse, Jakub M. Tomczak, Christos Louizos +1
stat.MLcs.LGarXiv:1905.10427v22019On the Origin of Implicit Regularization in Stochastic Gradient Descent
Samuel L. Smith, Benoit Dherin, David G. T. Barrett +1
cs.LGstat.MLarXiv:2101.12176v12021Adversarial Filters of Dataset Biases
Ronan Le Bras, Swabha Swayamdipta, Chandra Bhagavatula +4
cs.LGcs.AIcs.CLarXiv:2002.04108v32020Deep Gaussian Processes for Regression using Approximate Expectation Propagation
Thang D. Bui, Daniel Hernández-Lobato, Yingzhen Li +2
stat.MLcs.LGarXiv:1602.04133v12016Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees
Yuping Luo, Huazhe Xu, Yuanzhi Li +3
cs.LGcs.AIstat.MLarXiv:1807.03858v520182 OLMo 2 Furious
Team OLMo, Pete Walsh, Luca Soldaini +40
cs.CLcs.LGarXiv:2501.00656v32024The Unreasonable Ineffectiveness of the Deeper Layers
Andrey Gromov, Kushal Tirumala, Hassan Shapourian +2
cs.CLcs.LGstat.MLarXiv:2403.17887v22024Learning to cluster in order to transfer across domains and tasks
Yen-Chang Hsu, Zhaoyang Lv, Zsolt Kira
cs.LGcs.AIcs.CVarXiv:1711.10125v32017Progressive Prompts: Continual Learning for Language Models
Anastasia Razdaibiedina, Yuning Mao, Rui Hou +3
cs.CLcs.AIcs.LGarXiv:2301.12314v12023Defending Against Indirect Prompt Injection Attacks With Spotlighting
Keegan Hines, Gary Lopez, Matthew Hall +3
cs.CRcs.CLcs.LGarXiv:2403.14720v12024Autoregressive Diffusion Models
Emiel Hoogeboom, Alexey A. Gritsenko, Jasmijn Bastings +3
cs.LGstat.MLarXiv:2110.02037v22021FLamby: Datasets and Benchmarks for Cross-Silo Federated Learning in Realistic Healthcare Settings
Jean Ogier du Terrail, Samy-Safwan Ayed, Edwige Cyffers +21
cs.LGcs.CVarXiv:2210.04620v32022Combined Scaling for Zero-shot Transfer Learning
Hieu Pham, Zihang Dai, Golnaz Ghiasi +9
cs.LGcs.CLcs.CVarXiv:2111.10050v32021ViNT: A Foundation Model for Visual Navigation
Dhruv Shah, Ajay Sridhar, Nitish Dashora +4
cs.ROcs.CVcs.LGarXiv:2306.14846v22023Variable Impedance Control in End-Effector Space: An Action Space for Reinforcement Learning in Contact-Rich Tasks
Roberto Martín-Martín, Michelle A. Lee, Rachel Gardner +3
cs.ROcs.AIcs.LGarXiv:1906.08880v22019TACS: Trajectory-Aware Candidate Selection for LLM Jailbreak Suffix Optimization
Shiliang Xiao
cs.CLcs.LGarXiv:2608.29564v12026Field-weighted Factorization Machines for Click-Through Rate Prediction in Display Advertising
Junwei Pan, Jian Xu, Alfonso Lobos Ruiz +4
cs.LGstat.MLarXiv:1806.03514v22018Detection of Novel Social Bots by Ensembles of Specialized Classifiers
Mohsen Sayyadiharikandeh, Onur Varol, Kai-Cheng Yang +2
cs.SIcs.IRcs.LGarXiv:2006.06867v22020CodeNeRF: Disentangled Neural Radiance Fields for Object Categories
Wonbong Jang, Lourdes Agapito
cs.GRcs.CVcs.LGarXiv:2109.01750v12021Structured Pruning Learns Compact and Accurate Models
Mengzhou Xia, Zexuan Zhong, Danqi Chen
cs.CLcs.LGarXiv:2204.00408v32022LoGo: Token-Level Dynamic Local-Global Attention
Yuqi Pan, Zheng Li, Bohao Tang +2
cs.CLcs.LGarXiv:2608.29539v12026DarkneTZ: Towards Model Privacy at the Edge using Trusted Execution Environments
Fan Mo, Ali Shahin Shamsabadi, Kleomenis Katevas +4
cs.LGcs.CRstat.MLarXiv:2004.05703v12020DeepSight: Mitigating Backdoor Attacks in Federated Learning Through Deep Model Inspection
Phillip Rieger, Thien Duc Nguyen, Markus Miettinen +1
cs.CRcs.LGarXiv:2201.00763v12022Predicting Head Movement in Panoramic Video: A Deep Reinforcement Learning Approach
Yuhang Song, Mai Xu, Jianyi Wang +3
cs.CVcs.LGarXiv:1710.10755v52017QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving
Yujun Lin, Haotian Tang, Shang Yang +4
cs.CLcs.AIcs.LGarXiv:2405.04532v32024A Simple Convergence Proof of Adam and Adagrad
Alexandre Défossez, Léon Bottou, Francis Bach +1
stat.MLcs.LGarXiv:2003.02395v32020