Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
11,401 to 11,460 of 20,016
Towards Understanding Grokking: An Effective Theory of Representation Learning
Ziming Liu, Ouail Kitouni, Niklas Nolte +3
cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2205.10343v22022Inductive Matrix Completion Based on Graph Neural Networks
Muhan Zhang, Yixin Chen
cs.IRcs.LGstat.MLarXiv:1904.12058v32019Image Reconstruction: From Sparsity to Data-adaptive Methods and Machine Learning
Saiprasad Ravishankar, Jong Chul Ye, Jeffrey A. Fessler
eess.IVcs.LGstat.MLarXiv:1904.02816v32019Double Graph Based Reasoning for Document-level Relation Extraction
Shuang Zeng, Runxin Xu, Baobao Chang +1
cs.CLcs.AIcs.LGarXiv:2009.13752v12020What Can Neural Networks Reason About?
Keyulu Xu, Jingling Li, Mozhi Zhang +3
cs.LGcs.AIcs.CVarXiv:1905.13211v42019Geometric Latent Diffusion Models for 3D Molecule Generation
Minkai Xu, Alexander Powers, Ron Dror +2
cs.LGq-bio.QMarXiv:2305.01140v12023Training Neural Networks with Fixed Sparse Masks
Yi-Lin Sung, Varun Nair, Colin Raffel
cs.LGarXiv:2111.09839v12021Benchmarking Simulation-Based Inference
Jan-Matthis Lueckmann, Jan Boelts, David S. Greenberg +2
stat.MLcs.LGarXiv:2101.04653v22021Neural Symbolic Regression that Scales
Luca Biggio, Tommaso Bendinelli, Alexander Neitz +2
cs.LGarXiv:2106.06427v12021A Survey on Intelligent Internet of Things: Applications, Security, Privacy, and Future Directions
Ons Aouedi, Thai-Hoc Vu, Alessio Sacco +4
cs.NIcs.AIcs.CRarXiv:2406.03820v22024VERSE: Versatile Graph Embeddings from Similarity Measures
Anton Tsitsulin, Davide Mottin, Panagiotis Karras +1
cs.SIcs.LGarXiv:1803.04742v12018Adversarial Examples that Fool both Computer Vision and Time-Limited Humans
Gamaleldin F. Elsayed, Shreya Shankar, Brian Cheung +4
cs.LGcs.CVq-bio.NCarXiv:1802.08195v32018LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models
Yukang Chen, Shengju Qian, Haotian Tang +4
cs.CLcs.AIcs.LGarXiv:2309.12307v32023EEG-GAN: Generative adversarial networks for electroencephalograhic (EEG) brain signals
Kay Gregor Hartmann, Robin Tibor Schirrmeister, Tonio Ball
eess.SPcs.LGq-bio.NCarXiv:1806.01875v12018Fast, Accurate and Lightweight Super-Resolution with Neural Architecture Search
Xiangxiang Chu, Bo Zhang, Hailong Ma +2
cs.CVcs.LGarXiv:1901.07261v32019What the Constant Velocity Model Can Teach Us About Pedestrian Motion Prediction
Christoph Schöller, Vincent Aravantinos, Florian Lay +1
cs.CVcs.LGcs.ROarXiv:1903.07933v32019Var-CNN: A Data-Efficient Website Fingerprinting Attack Based on Deep Learning
Sanjit Bhat, David Lu, Albert Kwon +1
cs.CRcs.LGarXiv:1802.10215v22018Zero-Shot Entity Linking by Reading Entity Descriptions
Lajanugen Logeswaran, Ming-Wei Chang, Kenton Lee +3
cs.CLcs.LGarXiv:1906.07348v12019Differentiable Causal Discovery from Interventional Data
Philippe Brouillard, Sébastien Lachapelle, Alexandre Lacoste +2
cs.LGstat.MLarXiv:2007.01754v22020Zero-Shot Task Generalization with Multi-Task Deep Reinforcement Learning
Junhyuk Oh, Satinder Singh, Honglak Lee +1
cs.AIcs.LGarXiv:1706.05064v22017Combining data assimilation and machine learning to emulate a dynamical model from sparse and noisy observations: a case study with the Lorenz 96 model
Julien Brajard, Alberto Carassi, Marc Bocquet +1
stat.MLcs.LGphysics.ao-pharXiv:2001.01520v22020Multi-task Neural Networks for QSAR Predictions
George E. Dahl, Navdeep Jaitly, Ruslan Salakhutdinov
stat.MLcs.LGcs.NEarXiv:1406.1231v12014Deep Bidirectional Language-Knowledge Graph Pretraining
Michihiro Yasunaga, Antoine Bosselut, Hongyu Ren +4
cs.CLcs.AIcs.LGarXiv:2210.09338v22022PeCo: Perceptual Codebook for BERT Pre-training of Vision Transformers
Xiaoyi Dong, Jianmin Bao, Ting Zhang +7
cs.CVcs.LGarXiv:2111.12710v32021Deep $k$-Means: Jointly clustering with $k$-Means and learning representations
Maziar Moradi Fard, Thibaut Thonet, Eric Gaussier
cs.LGstat.MLarXiv:1806.10069v22018RetainVis: Visual Analytics with Interpretable and Interactive Recurrent Neural Networks on Electronic Medical Records
Bum Chul Kwon, Min-Je Choi, Joanne Taery Kim +5
cs.LGcs.HCstat.MLarXiv:1805.10724v32018Transfer Learning for Non-Intrusive Load Monitoring
Michele DIncecco, Stefano Squartini, Mingjun Zhong
cs.LGstat.MLarXiv:1902.08835v32019Transformers Learn Shortcuts to Automata
Bingbin Liu, Jordan T. Ash, Surbhi Goel +2
cs.LGcs.FLstat.MLarXiv:2210.10749v22022Deep Convolutional Networks as shallow Gaussian Processes
Adrià Garriga-Alonso, Carl Edward Rasmussen, Laurence Aitchison
stat.MLcs.LGarXiv:1808.05587v22018Meta-GNN: On Few-shot Node Classification in Graph Meta-learning
Fan Zhou, Chengtai Cao, Kunpeng Zhang +3
cs.LGstat.MLarXiv:1905.09718v12019Stronger Data Poisoning Attacks Break Data Sanitization Defenses
Pang Wei Koh, Jacob Steinhardt, Percy Liang
stat.MLcs.CRcs.LGarXiv:1811.00741v22018What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Marcin Andrychowicz, Anton Raichuk, Piotr Stańczyk +9
cs.LGstat.MLarXiv:2006.05990v12020Riemannian Score-Based Generative Modelling
Valentin De Bortoli, Emile Mathieu, Michael Hutchinson +3
cs.LGmath.PRstat.MLarXiv:2202.02763v32022Can Bad Teaching Induce Forgetting? Unlearning in Deep Networks using an Incompetent Teacher
Vikram S Chundawat, Ayush K Tarun, Murari Mandal +1
cs.LGarXiv:2205.08096v22022Vision-Language Models for Medical Report Generation and Visual Question Answering: A Review
Iryna Hartsock, Ghulam Rasool
cs.CVcs.LGarXiv:2403.02469v22024Foundations and Trends in Multimodal Machine Learning: Principles, Challenges, and Open Questions
Paul Pu Liang, Amir Zadeh, Louis-Philippe Morency
cs.LGcs.AIcs.CLarXiv:2209.03430v22022Jasper: An End-to-End Convolutional Neural Acoustic Model
Jason Li, Vitaly Lavrukhin, Boris Ginsburg +5
eess.AScs.CLcs.LGarXiv:1904.03288v32019Efficiently Trainable Text-to-Speech System Based on Deep Convolutional Networks with Guided Attention
Hideyuki Tachibana, Katsuya Uenoyama, Shunsuke Aihara
cs.SDcs.AIcs.LGarXiv:1710.08969v22017Relational Pooling for Graph Representations
Ryan L. Murphy, Balasubramaniam Srinivasan, Vinayak Rao +1
cs.LGstat.MLarXiv:1903.02541v22019Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
Tuomas Haarnoja, Ben Moran, Guy Lever +25
cs.ROcs.AIcs.LGarXiv:2304.13653v22023Conditional Image Generation with Score-Based Diffusion Models
Georgios Batzolis, Jan Stanczuk, Carola-Bibiane Schönlieb +1
cs.LGcs.CVstat.MLarXiv:2111.13606v12021Ridge Regression, Hubness, and Zero-Shot Learning
Yutaro Shigeto, Ikumi Suzuki, Kazuo Hara +2
cs.LGstat.MLarXiv:1507.00825v12015Reachability Analysis of Deep Neural Networks with Provable Guarantees
Wenjie Ruan, Xiaowei Huang, Marta Kwiatkowska
cs.LGcs.CVstat.MLarXiv:1805.02242v12018CrossCodeEval: A Diverse and Multilingual Benchmark for Cross-File Code Completion
Yangruibo Ding, Zijian Wang, Wasi Uddin Ahmad +8
cs.LGcs.CLcs.SEarXiv:2310.11248v22023Unsupervised Semantic Segmentation by Contrasting Object Mask Proposals
Wouter Van Gansbeke, Simon Vandenhende, Stamatios Georgoulis +1
cs.CVcs.LGarXiv:2102.06191v32021TweepFake: about Detecting Deepfake Tweets
Tiziano Fagni, Fabrizio Falchi, Margherita Gambini +2
cs.CLcs.LGarXiv:2008.00036v22020Gated-Attention Architectures for Task-Oriented Language Grounding
Devendra Singh Chaplot, Kanthashree Mysore Sathyendra, Rama Kumar Pasumarthi +2
cs.LGcs.AIcs.CLarXiv:1706.07230v22017Text-based Editing of Talking-head Video
Ohad Fried, Ayush Tewari, Michael Zollhöfer +7
cs.CVcs.GRcs.LGarXiv:1906.01524v12019The Lyapunov Neural Network: Adaptive Stability Certification for Safe Learning of Dynamical Systems
Spencer M. Richards, Felix Berkenkamp, Andreas Krause
eess.SYcs.LGcs.ROarXiv:1808.00924v22018A Boundary Tilting Persepective on the Phenomenon of Adversarial Examples
Thomas Tanay, Lewis Griffin
cs.LGstat.MLarXiv:1608.07690v12016AutoGAN: Neural Architecture Search for Generative Adversarial Networks
Xinyu Gong, Shiyu Chang, Yifan Jiang +1
cs.CVcs.LGeess.IVarXiv:1908.03835v12019AttentionGAN: Unpaired Image-to-Image Translation using Attention-Guided Generative Adversarial Networks
Hao Tang, Hong Liu, Dan Xu +2
cs.CVcs.LGeess.IVarXiv:1911.11897v52019Score-Based Generative Modeling with Critically-Damped Langevin Diffusion
Tim Dockhorn, Arash Vahdat, Karsten Kreis
stat.MLcs.LGarXiv:2112.07068v42021Fast and Complete: Enabling Complete Neural Network Verification with Rapid and Massively Parallel Incomplete Verifiers
Kaidi Xu, Huan Zhang, Shiqi Wang +4
cs.AIcs.LGarXiv:2011.13824v22020Scalable Multi-Hop Relational Reasoning for Knowledge-Aware Question Answering
Yanlin Feng, Xinyue Chen, Bill Yuchen Lin +3
cs.CLcs.LGarXiv:2005.00646v22020On the Role of Sparsity and DAG Constraints for Learning Linear DAGs
Ignavier Ng, AmirEmad Ghassami, Kun Zhang
cs.LGstat.MLarXiv:2006.10201v32020Pre-Training with Whole Word Masking for Chinese BERT
Yiming Cui, Wanxiang Che, Ting Liu +2
cs.CLcs.LGarXiv:1906.08101v32019Hyperparameter Importance Across Datasets
J. N. van Rijn, F. Hutter
stat.MLcs.LGarXiv:1710.04725v22017A review on distance based time series classification
Amaia Abanda, Usue Mori, Jose A. Lozano
stat.MLcs.LGarXiv:1806.04509v12018End-to-End Robotic Reinforcement Learning without Reward Engineering
Avi Singh, Larry Yang, Kristian Hartikainen +2
cs.LGcs.CVcs.ROarXiv:1904.07854v22019