Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

13,261 to 13,320 of 20,199

  1. Pyramidal Flow Matching for Efficient Video Generative Modeling

    Yang Jin, Zhicheng Sun, Ningyuan Li +8

    cs.CVcs.LGarXiv:2410.05954v22024
  2. Loss landscapes and optimization in over-parameterized non-linear systems and neural networks

    Chaoyue Liu, Libin Zhu, Mikhail Belkin

    cs.LGmath.OCstat.MLarXiv:2003.00307v22020
  3. Natural Language Inference by Tree-Based Convolution and Heuristic Matching

    Lili Mou, Rui Men, Ge Li +4

    cs.CLcs.LGarXiv:1512.08422v32015
  4. Semantic Photo Manipulation with a Generative Image Prior

    David Bau, Hendrik Strobelt, William Peebles +4

    cs.CVcs.GRcs.LGarXiv:2005.07727v22020
  5. Physics-Informed Neural Networks for Multiphysics Data Assimilation with Application to Subsurface Transport

    QiZhi He, David Brajas-Solano, Guzel Tartakovsky +1

    cs.LGphysics.comp-phstat.MLarXiv:1912.02968v12019
  6. The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models

    Alexander Pan, Kush Bhatia, Jacob Steinhardt

    cs.LGcs.AIstat.MLarXiv:2201.03544v22022
  7. Preparing for the Unknown: Learning a Universal Policy with Online System Identification

    Wenhao Yu, Jie Tan, C. Karen Liu +1

    cs.LGcs.ROeess.SYarXiv:1702.02453v32017
  8. Efficient and Modular Implicit Differentiation

    Mathieu Blondel, Quentin Berthet, Marco Cuturi +5

    cs.LGmath.NAstat.MLarXiv:2105.15183v52021
  9. A Survey on Generative Adversarial Networks: Variants, Applications, and Training

    Abdul Jabbar, Xi Li, Bourahla Omar

    cs.CVcs.LGeess.IVarXiv:2006.05132v12020
  10. Ablating Concepts in Text-to-Image Diffusion Models

    Nupur Kumari, Bingliang Zhang, Sheng-Yu Wang +3

    cs.CVcs.GRcs.LGarXiv:2303.13516v32023
  11. Android in the Wild: A Large-Scale Dataset for Android Device Control

    Christopher Rawles, Alice Li, Daniel Rodriguez +2

    cs.LGcs.CLcs.HCarXiv:2307.10088v22023
  12. A comparison of deep networks with ReLU activation function and linear spline-type methods

    Konstantin Eckle, Johannes Schmidt-Hieber

    stat.MLcs.LGstat.MEarXiv:1804.02253v22018
  13. SQA3D: Situated Question Answering in 3D Scenes

    Xiaojian Ma, Silong Yong, Zilong Zheng +4

    cs.CVcs.AIcs.CLarXiv:2210.07474v52022
  14. Poisoning Language Models During Instruction Tuning

    Alexander Wan, Eric Wallace, Sheng Shen +1

    cs.CLcs.CRcs.LGarXiv:2305.00944v12023
  15. Language Models Represent Space and Time

    Wes Gurnee, Max Tegmark

    cs.LGcs.AIcs.CLarXiv:2310.02207v32023
  16. Recurrent Independent Mechanisms

    Anirudh Goyal, Alex Lamb, Jordan Hoffmann +4

    cs.LGcs.AIstat.MLarXiv:1909.10893v62019
  17. Sequence-to-Sequence Models Can Directly Translate Foreign Speech

    Ron J. Weiss, Jan Chorowski, Navdeep Jaitly +2

    cs.CLcs.LGstat.MLarXiv:1703.08581v22017
  18. TensorFuzz: Debugging Neural Networks with Coverage-Guided Fuzzing

    Augustus Odena, Ian Goodfellow

    stat.MLcs.LGarXiv:1807.10875v12018
  19. Transformer Quality in Linear Time

    Weizhe Hua, Zihang Dai, Hanxiao Liu +1

    cs.LGcs.AIcs.CLarXiv:2202.10447v22022
  20. TensorFlow-Serving: Flexible, High-Performance ML Serving

    Christopher Olston, Noah Fiedel, Kiril Gorovoy +6

    cs.DCcs.LGarXiv:1712.06139v22017
  21. "Double-DIP": Unsupervised Image Decomposition via Coupled Deep-Image-Priors

    Yossi Gandelsman, Assaf Shocher, Michal Irani

    cs.CVcs.LGarXiv:1812.00467v22018
  22. FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture Search

    Xiangxiang Chu, Bo Zhang, Ruijun Xu

    cs.LGcs.AIcs.CVarXiv:1907.01845v52019
  23. Trained Transformers Learn Linear Models In-Context

    Ruiqi Zhang, Spencer Frei, Peter L. Bartlett

    stat.MLcs.AIcs.CLarXiv:2306.09927v32023
  24. History Repeats Itself: Human Motion Prediction via Motion Attention

    Wei Mao, Miaomiao Liu, Mathieu Salzmann

    cs.CVcs.LGeess.IVarXiv:2007.11755v12020
  25. Improving performance of CNN to predict likelihood of COVID-19 using chest X-ray images with preprocessing algorithms

    Morteza Heidari, Seyedehnafiseh Mirniaharikandehei, Abolfazl Zargari Khuzani +3

    eess.IVcs.LGarXiv:2006.12229v12020
  26. Additive Gaussian Processes

    David Duvenaud, Hannes Nickisch, Carl Edward Rasmussen

    stat.MLcs.LGarXiv:1112.4394v12011
  27. How to train your neural ODE: the world of Jacobian and kinetic regularization

    Chris Finlay, Jörn-Henrik Jacobsen, Levon Nurbekyan +1

    stat.MLcs.LGarXiv:2002.02798v32020
  28. DPatch: An Adversarial Patch Attack on Object Detectors

    Xin Liu, Huanrui Yang, Ziwei Liu +3

    cs.CVcs.CRcs.LGarXiv:1806.02299v42018
  29. SPINN: Synergistic Progressive Inference of Neural Networks over Device and Cloud

    Stefanos Laskaridis, Stylianos I. Venieris, Mario Almeida +2

    cs.LGcs.CVcs.DCarXiv:2008.06402v22020
  30. Multi-level Feature Learning for Contrastive Multi-view Clustering

    Jie Xu, Huayi Tang, Yazhou Ren +3

    cs.LGcs.CVarXiv:2106.11193v22021
  31. Minimal Gated Unit for Recurrent Neural Networks

    Guo-Bing Zhou, Jianxin Wu, Chen-Lin Zhang +1

    cs.NEcs.LGarXiv:1603.09420v12016
  32. DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism

    Jinglin Liu, Chengxi Li, Yi Ren +2

    eess.AScs.LGcs.SDarXiv:2105.02446v62021
  33. Deep Learning for Unsupervised Insider Threat Detection in Structured Cybersecurity Data Streams

    Aaron Tuor, Samuel Kaplan, Brian Hutchinson +2

    cs.NEcs.CRcs.LGarXiv:1710.00811v22017
  34. Combating Adversarial Misspellings with Robust Word Recognition

    Danish Pruthi, Bhuwan Dhingra, Zachary C. Lipton

    cs.CLcs.CRcs.LGarXiv:1905.11268v22019
  35. Multi-Time Attention Networks for Irregularly Sampled Time Series

    Satya Narayan Shukla, Benjamin M. Marlin

    cs.LGcs.AIarXiv:2101.10318v22021
  36. EAGLE-2: Faster Inference of Language Models with Dynamic Draft Trees

    Yuhui Li, Fangyun Wei, Chao Zhang +1

    cs.CLcs.LGarXiv:2406.16858v22024
  37. Streaming Variational Bayes

    Tamara Broderick, Nicholas Boyd, Andre Wibisono +2

    stat.MLcs.LGarXiv:1307.6769v22013
  38. Real-Time Intermediate Flow Estimation for Video Frame Interpolation

    Zhewei Huang, Tianyuan Zhang, Wen Heng +2

    cs.CVcs.LGarXiv:2011.06294v122020
  39. LCSTS: A Large Scale Chinese Short Text Summarization Dataset

    Baotian Hu, Qingcai Chen, Fangze Zhu

    cs.CLcs.IRcs.LGarXiv:1506.05865v42015
  40. Regularization for Deep Learning: A Taxonomy

    Jan Kukačka, Vladimir Golkov, Daniel Cremers

    cs.LGcs.AIcs.CVarXiv:1710.10686v12017
  41. Emergent Linear Representations in World Models of Self-Supervised Sequence Models

    Neel Nanda, Andrew Lee, Martin Wattenberg

    cs.LGarXiv:2309.00941v22023
  42. Inverting The Generator Of A Generative Adversarial Network

    Antonia Creswell, Anil Anthony Bharath

    cs.CVcs.LGarXiv:1611.05644v12016
  43. Being Bayesian, Even Just a Bit, Fixes Overconfidence in ReLU Networks

    Agustinus Kristiadi, Matthias Hein, Philipp Hennig

    stat.MLcs.LGarXiv:2002.10118v22020
  44. Deceiving Google's Perspective API Built for Detecting Toxic Comments

    Hossein Hosseini, Sreeram Kannan, Baosen Zhang +1

    cs.LGcs.CYcs.SIarXiv:1702.08138v12017
  45. InceptionNeXt: When Inception Meets ConvNeXt

    Weihao Yu, Pan Zhou, Shuicheng Yan +1

    cs.CVcs.AIcs.LGarXiv:2303.16900v32023
  46. Low-Resource Languages Jailbreak GPT-4

    Zheng-Xin Yong, Cristina Menghini, Stephen H. Bach

    cs.CLcs.AIcs.CRarXiv:2310.02446v22023
  47. Extreme Parkour with Legged Robots

    Xuxin Cheng, Kexin Shi, Ananye Agarwal +1

    cs.ROcs.AIcs.CVarXiv:2309.14341v12023
  48. Multimodal Virtual Point 3D Detection

    Tianwei Yin, Xingyi Zhou, Philipp Krähenbühl

    cs.CVcs.LGcs.ROarXiv:2111.06881v12021
  49. Generalization and Representational Limits of Graph Neural Networks

    Vikas K. Garg, Stefanie Jegelka, Tommi Jaakkola

    cs.LGstat.MLarXiv:2002.06157v12020
  50. Learning to Optimize: A Primer and A Benchmark

    Tianlong Chen, Xiaohan Chen, Wuyang Chen +4

    math.OCcs.LGstat.MLarXiv:2103.12828v22021
  51. Optimal Transport for structured data with application on graphs

    Titouan Vayer, Laetitia Chapel, Rémi Flamary +2

    stat.MLcs.LGarXiv:1805.09114v32018
    Summaries:한국어
  52. Graph Embedding on Biomedical Networks: Methods, Applications, and Evaluations

    Xiang Yue, Zhen Wang, Jingong Huang +7

    cs.LGcs.SIarXiv:1906.05017v32019
  53. Unsupervised learning of phase transitions: from principal component analysis to variational autoencoders

    Sebastian Johann Wetzel

    cond-mat.stat-mechcs.LGstat.MLarXiv:1703.02435v22017
  54. Skip Connections Matter: On the Transferability of Adversarial Examples Generated with ResNets

    Dongxian Wu, Yisen Wang, Shu-Tao Xia +2

    cs.LGcs.CRcs.CVarXiv:2002.05990v12020
  55. Deep Convolutional Neural Networks for Raman Spectrum Recognition: A Unified Solution

    Jinchao Liu, Margarita Osadchy, Lorna Ashton +3

    cs.LGstat.MLarXiv:1708.09022v12017
    Summaries:한국어
  56. CUAD: An Expert-Annotated NLP Dataset for Legal Contract Review

    Dan Hendrycks, Collin Burns, Anya Chen +1

    cs.CLcs.LGarXiv:2103.06268v22021
  57. Utterance-level Aggregation For Speaker Recognition In The Wild

    Weidi Xie, Arsha Nagrani, Joon Son Chung +1

    eess.AScs.LGcs.MMarXiv:1902.10107v22019
  58. 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations

    Tsung-Wei Ke, Nikolaos Gkanatsios, Katerina Fragkiadaki

    cs.ROcs.AIcs.CVarXiv:2402.10885v32024
  59. Learning how to explain neural networks: PatternNet and PatternAttribution

    Pieter-Jan Kindermans, Kristof T. Schütt, Maximilian Alber +4

    stat.MLcs.LGarXiv:1705.05598v22017
  60. Federated Learning with Cooperating Devices: A Consensus Approach for Massive IoT Networks

    Stefano Savazzi, Monica Nicoli, Vittorio Rampa

    eess.SPcs.DCcs.LGarXiv:1912.13163v12019