Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
13,261 to 13,320 of 20,199
Pyramidal Flow Matching for Efficient Video Generative Modeling
Yang Jin, Zhicheng Sun, Ningyuan Li +8
cs.CVcs.LGarXiv:2410.05954v22024Loss landscapes and optimization in over-parameterized non-linear systems and neural networks
Chaoyue Liu, Libin Zhu, Mikhail Belkin
cs.LGmath.OCstat.MLarXiv:2003.00307v22020Natural Language Inference by Tree-Based Convolution and Heuristic Matching
Lili Mou, Rui Men, Ge Li +4
cs.CLcs.LGarXiv:1512.08422v32015Semantic Photo Manipulation with a Generative Image Prior
David Bau, Hendrik Strobelt, William Peebles +4
cs.CVcs.GRcs.LGarXiv:2005.07727v22020Physics-Informed Neural Networks for Multiphysics Data Assimilation with Application to Subsurface Transport
QiZhi He, David Brajas-Solano, Guzel Tartakovsky +1
cs.LGphysics.comp-phstat.MLarXiv:1912.02968v12019The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models
Alexander Pan, Kush Bhatia, Jacob Steinhardt
cs.LGcs.AIstat.MLarXiv:2201.03544v22022Preparing for the Unknown: Learning a Universal Policy with Online System Identification
Wenhao Yu, Jie Tan, C. Karen Liu +1
cs.LGcs.ROeess.SYarXiv:1702.02453v32017Efficient and Modular Implicit Differentiation
Mathieu Blondel, Quentin Berthet, Marco Cuturi +5
cs.LGmath.NAstat.MLarXiv:2105.15183v52021A Survey on Generative Adversarial Networks: Variants, Applications, and Training
Abdul Jabbar, Xi Li, Bourahla Omar
cs.CVcs.LGeess.IVarXiv:2006.05132v12020Ablating Concepts in Text-to-Image Diffusion Models
Nupur Kumari, Bingliang Zhang, Sheng-Yu Wang +3
cs.CVcs.GRcs.LGarXiv:2303.13516v32023Android in the Wild: A Large-Scale Dataset for Android Device Control
Christopher Rawles, Alice Li, Daniel Rodriguez +2
cs.LGcs.CLcs.HCarXiv:2307.10088v22023A comparison of deep networks with ReLU activation function and linear spline-type methods
Konstantin Eckle, Johannes Schmidt-Hieber
stat.MLcs.LGstat.MEarXiv:1804.02253v22018SQA3D: Situated Question Answering in 3D Scenes
Xiaojian Ma, Silong Yong, Zilong Zheng +4
cs.CVcs.AIcs.CLarXiv:2210.07474v52022Poisoning Language Models During Instruction Tuning
Alexander Wan, Eric Wallace, Sheng Shen +1
cs.CLcs.CRcs.LGarXiv:2305.00944v12023Language Models Represent Space and Time
Wes Gurnee, Max Tegmark
cs.LGcs.AIcs.CLarXiv:2310.02207v32023Recurrent Independent Mechanisms
Anirudh Goyal, Alex Lamb, Jordan Hoffmann +4
cs.LGcs.AIstat.MLarXiv:1909.10893v62019Sequence-to-Sequence Models Can Directly Translate Foreign Speech
Ron J. Weiss, Jan Chorowski, Navdeep Jaitly +2
cs.CLcs.LGstat.MLarXiv:1703.08581v22017TensorFuzz: Debugging Neural Networks with Coverage-Guided Fuzzing
Augustus Odena, Ian Goodfellow
stat.MLcs.LGarXiv:1807.10875v12018Transformer Quality in Linear Time
Weizhe Hua, Zihang Dai, Hanxiao Liu +1
cs.LGcs.AIcs.CLarXiv:2202.10447v22022TensorFlow-Serving: Flexible, High-Performance ML Serving
Christopher Olston, Noah Fiedel, Kiril Gorovoy +6
cs.DCcs.LGarXiv:1712.06139v22017"Double-DIP": Unsupervised Image Decomposition via Coupled Deep-Image-Priors
Yossi Gandelsman, Assaf Shocher, Michal Irani
cs.CVcs.LGarXiv:1812.00467v22018FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture Search
Xiangxiang Chu, Bo Zhang, Ruijun Xu
cs.LGcs.AIcs.CVarXiv:1907.01845v52019Trained Transformers Learn Linear Models In-Context
Ruiqi Zhang, Spencer Frei, Peter L. Bartlett
stat.MLcs.AIcs.CLarXiv:2306.09927v32023History Repeats Itself: Human Motion Prediction via Motion Attention
Wei Mao, Miaomiao Liu, Mathieu Salzmann
cs.CVcs.LGeess.IVarXiv:2007.11755v12020Improving performance of CNN to predict likelihood of COVID-19 using chest X-ray images with preprocessing algorithms
Morteza Heidari, Seyedehnafiseh Mirniaharikandehei, Abolfazl Zargari Khuzani +3
eess.IVcs.LGarXiv:2006.12229v12020Additive Gaussian Processes
David Duvenaud, Hannes Nickisch, Carl Edward Rasmussen
stat.MLcs.LGarXiv:1112.4394v12011How to train your neural ODE: the world of Jacobian and kinetic regularization
Chris Finlay, Jörn-Henrik Jacobsen, Levon Nurbekyan +1
stat.MLcs.LGarXiv:2002.02798v32020DPatch: An Adversarial Patch Attack on Object Detectors
Xin Liu, Huanrui Yang, Ziwei Liu +3
cs.CVcs.CRcs.LGarXiv:1806.02299v42018SPINN: Synergistic Progressive Inference of Neural Networks over Device and Cloud
Stefanos Laskaridis, Stylianos I. Venieris, Mario Almeida +2
cs.LGcs.CVcs.DCarXiv:2008.06402v22020Multi-level Feature Learning for Contrastive Multi-view Clustering
Jie Xu, Huayi Tang, Yazhou Ren +3
cs.LGcs.CVarXiv:2106.11193v22021Minimal Gated Unit for Recurrent Neural Networks
Guo-Bing Zhou, Jianxin Wu, Chen-Lin Zhang +1
cs.NEcs.LGarXiv:1603.09420v12016DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism
Jinglin Liu, Chengxi Li, Yi Ren +2
eess.AScs.LGcs.SDarXiv:2105.02446v62021Deep Learning for Unsupervised Insider Threat Detection in Structured Cybersecurity Data Streams
Aaron Tuor, Samuel Kaplan, Brian Hutchinson +2
cs.NEcs.CRcs.LGarXiv:1710.00811v22017Combating Adversarial Misspellings with Robust Word Recognition
Danish Pruthi, Bhuwan Dhingra, Zachary C. Lipton
cs.CLcs.CRcs.LGarXiv:1905.11268v22019Multi-Time Attention Networks for Irregularly Sampled Time Series
Satya Narayan Shukla, Benjamin M. Marlin
cs.LGcs.AIarXiv:2101.10318v22021EAGLE-2: Faster Inference of Language Models with Dynamic Draft Trees
Yuhui Li, Fangyun Wei, Chao Zhang +1
cs.CLcs.LGarXiv:2406.16858v22024Streaming Variational Bayes
Tamara Broderick, Nicholas Boyd, Andre Wibisono +2
stat.MLcs.LGarXiv:1307.6769v22013Real-Time Intermediate Flow Estimation for Video Frame Interpolation
Zhewei Huang, Tianyuan Zhang, Wen Heng +2
cs.CVcs.LGarXiv:2011.06294v122020LCSTS: A Large Scale Chinese Short Text Summarization Dataset
Baotian Hu, Qingcai Chen, Fangze Zhu
cs.CLcs.IRcs.LGarXiv:1506.05865v42015Regularization for Deep Learning: A Taxonomy
Jan Kukačka, Vladimir Golkov, Daniel Cremers
cs.LGcs.AIcs.CVarXiv:1710.10686v12017Emergent Linear Representations in World Models of Self-Supervised Sequence Models
Neel Nanda, Andrew Lee, Martin Wattenberg
cs.LGarXiv:2309.00941v22023Inverting The Generator Of A Generative Adversarial Network
Antonia Creswell, Anil Anthony Bharath
cs.CVcs.LGarXiv:1611.05644v12016Being Bayesian, Even Just a Bit, Fixes Overconfidence in ReLU Networks
Agustinus Kristiadi, Matthias Hein, Philipp Hennig
stat.MLcs.LGarXiv:2002.10118v22020Deceiving Google's Perspective API Built for Detecting Toxic Comments
Hossein Hosseini, Sreeram Kannan, Baosen Zhang +1
cs.LGcs.CYcs.SIarXiv:1702.08138v12017InceptionNeXt: When Inception Meets ConvNeXt
Weihao Yu, Pan Zhou, Shuicheng Yan +1
cs.CVcs.AIcs.LGarXiv:2303.16900v32023Low-Resource Languages Jailbreak GPT-4
Zheng-Xin Yong, Cristina Menghini, Stephen H. Bach
cs.CLcs.AIcs.CRarXiv:2310.02446v22023Extreme Parkour with Legged Robots
Xuxin Cheng, Kexin Shi, Ananye Agarwal +1
cs.ROcs.AIcs.CVarXiv:2309.14341v12023Multimodal Virtual Point 3D Detection
Tianwei Yin, Xingyi Zhou, Philipp Krähenbühl
cs.CVcs.LGcs.ROarXiv:2111.06881v12021Generalization and Representational Limits of Graph Neural Networks
Vikas K. Garg, Stefanie Jegelka, Tommi Jaakkola
cs.LGstat.MLarXiv:2002.06157v12020Learning to Optimize: A Primer and A Benchmark
Tianlong Chen, Xiaohan Chen, Wuyang Chen +4
math.OCcs.LGstat.MLarXiv:2103.12828v22021Optimal Transport for structured data with application on graphs
Titouan Vayer, Laetitia Chapel, Rémi Flamary +2
stat.MLcs.LGarXiv:1805.09114v32018Summaries:한국어Graph Embedding on Biomedical Networks: Methods, Applications, and Evaluations
Xiang Yue, Zhen Wang, Jingong Huang +7
cs.LGcs.SIarXiv:1906.05017v32019Unsupervised learning of phase transitions: from principal component analysis to variational autoencoders
Sebastian Johann Wetzel
cond-mat.stat-mechcs.LGstat.MLarXiv:1703.02435v22017Skip Connections Matter: On the Transferability of Adversarial Examples Generated with ResNets
Dongxian Wu, Yisen Wang, Shu-Tao Xia +2
cs.LGcs.CRcs.CVarXiv:2002.05990v12020Deep Convolutional Neural Networks for Raman Spectrum Recognition: A Unified Solution
Jinchao Liu, Margarita Osadchy, Lorna Ashton +3
cs.LGstat.MLarXiv:1708.09022v12017Summaries:한국어CUAD: An Expert-Annotated NLP Dataset for Legal Contract Review
Dan Hendrycks, Collin Burns, Anya Chen +1
cs.CLcs.LGarXiv:2103.06268v22021Utterance-level Aggregation For Speaker Recognition In The Wild
Weidi Xie, Arsha Nagrani, Joon Son Chung +1
eess.AScs.LGcs.MMarXiv:1902.10107v220193D Diffuser Actor: Policy Diffusion with 3D Scene Representations
Tsung-Wei Ke, Nikolaos Gkanatsios, Katerina Fragkiadaki
cs.ROcs.AIcs.CVarXiv:2402.10885v32024Learning how to explain neural networks: PatternNet and PatternAttribution
Pieter-Jan Kindermans, Kristof T. Schütt, Maximilian Alber +4
stat.MLcs.LGarXiv:1705.05598v22017Federated Learning with Cooperating Devices: A Consensus Approach for Massive IoT Networks
Stefano Savazzi, Monica Nicoli, Vittorio Rampa
eess.SPcs.DCcs.LGarXiv:1912.13163v12019