Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

181 to 240 of 19,960

  1. NSGANetV2: Evolutionary Multi-Objective Surrogate-Assisted Neural Architecture Search

    Zhichao Lu, Kalyanmoy Deb, Erik Goodman +2

    cs.CVcs.LGcs.NEarXiv:2007.10396v12020
  2. Distribution Aligning Refinery of Pseudo-label for Imbalanced Semi-supervised Learning

    Jaehyung Kim, Youngbum Hur, Sejun Park +3

    cs.LGstat.MLarXiv:2007.08844v22020
  3. CardioState-JEPA: Delay-Aware Cross-Modal Learning of a Shared Cardiac Representation

    Hamza Shafiq, Hung Manh Pham, Bin Zhu +3

    cs.LGeess.IVstat.MLarXiv:2608.12944v12026
  4. Hands-on Bayesian Neural Networks -- a Tutorial for Deep Learning Users

    Laurent Valentin Jospin, Wray Buntine, Farid Boussaid +2

    cs.LGstat.MLarXiv:2007.06823v32020
  5. Multiscale Simulations of Complex Systems by Learning their Effective Dynamics

    Pantelis R. Vlachas, Georgios Arampatzis, Caroline Uhler +1

    physics.comp-phcs.LGnlin.CDarXiv:2006.13431v32020
  6. A Bayesian Approach to Robust Inverse Reinforcement Learning

    Ran Wei, Siliang Zeng, Chenliang Li +3

    cs.LGarXiv:2309.08571v22023
  7. TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs

    Yuxiang Zhang, Zhengxu Yu, Weihang Pan +5

    cs.LGcs.AIarXiv:2511.13223v12025
  8. TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools

    Shanghua Gao, Richard Zhu, Zhenglun Kong +5

    cs.AIcs.LGarXiv:2503.10970v12025
  9. Byzantine-Robust Learning on Heterogeneous Datasets via Bucketing

    Sai Praneeth Karimireddy, Lie He, Martin Jaggi

    cs.LGstat.MLarXiv:2006.09365v62020
  10. Training Generative Adversarial Networks with Limited Data

    Tero Karras, Miika Aittala, Janne Hellsten +3

    cs.CVcs.LGcs.NEarXiv:2006.06676v22020
  11. Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities

    Sreyan Ghosh, Zhifeng Kong, Sonal Kumar +6

    cs.SDcs.CLcs.LGarXiv:2503.03983v12025
  12. Deep Learning is Not So Mysterious or Different

    Andrew Gordon Wilson

    cs.LGstat.MLarXiv:2503.02113v22025
  13. Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks

    Patrick Lewis, Ethan Perez, Aleksandra Piktus +9

    cs.CLcs.LGarXiv:2005.11401v42020
    Summaries:한국어
  14. Dataset Distillation with Neural Characteristic Function: A Minmax Perspective

    Shaobo Wang, Yicun Yang, Zhiyuan Liu +4

    cs.CVcs.AIcs.LGarXiv:2502.20653v12025
    Summaries:한국어
  15. Universal Sparse Autoencoders: Interpretable Cross-Model Concept Alignment

    Harrish Thasarathan, Julian Forsyth, Thomas Fel +2

    cs.CVcs.LGarXiv:2502.03714v22025
  16. Maximum Density Divergence for Domain Adaptation

    Li Jingjing, Chen Erpeng, Ding Zhengming +3

    cs.CVcs.LGcs.MMarXiv:2004.12615v12020
  17. Reasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reasoning Models

    Junlin Fang, Do Nguyen-Thanh, Xiaogang Xu +2

    cs.AIcs.LGarXiv:2607.22098v12026
  18. Don't Judge an Object by Its Context: Learning to Overcome Contextual Bias

    Krishna Kumar Singh, Dhruv Mahajan, Kristen Grauman +3

    cs.CVcs.LGarXiv:2001.03152v22020
  19. Improving Medical Large Vision-Language Models with Abnormal-Aware Feedback

    Yucheng Zhou, Lingran Song, Jianbing Shen

    cs.CLcs.AIcs.CVarXiv:2501.01377v22025
  20. Frontier Models are Capable of In-context Scheming

    Alexander Meinke, Bronson Schoen, Jérémy Scheurer +3

    cs.AIcs.LGarXiv:2412.04984v22024
  21. Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training

    Zijian Zhang, Rizhen Hu, Athanasios Glentis +4

    cs.LGcs.CLarXiv:2607.01232v22026
  22. Omni-Scale CNNs: a simple and effective kernel size configuration for time series classification

    Wensi Tang, Guodong Long, Lu Liu +3

    cs.LGstat.MLarXiv:2002.10061v32020
  23. Beyond IID: How General Are Tabular Foundation Models, Really?

    Lennart Purucker, Andrej Tschalzev, Nick Erickson +7

    cs.LGcs.AIarXiv:2606.30410v12026
    Summaries:한국어
  24. Explaining Explanations: Axiomatic Feature Interactions for Deep Networks

    Joseph D. Janizek, Pascal Sturmfels, Su-In Lee

    cs.LGstat.MLarXiv:2002.04138v32020
  25. PhysGen: Rigid-Body Physics-Grounded Image-to-Video Generation

    Shaowei Liu, Zhongzheng Ren, Saurabh Gupta +1

    cs.CVcs.AIcs.LGarXiv:2409.18964v12024
  26. LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

    Yijia Xiao, Edward Sun, Tianyu Liu +1

    cs.AIcs.CLcs.CVarXiv:2407.04973v12024
  27. Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs

    Xin Lai, Zhuotao Tian, Yukang Chen +3

    cs.LGcs.AIcs.CLarXiv:2406.18629v12024
  28. SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

    John Yang, Carlos E. Jimenez, Alexander Wettig +4

    cs.SEcs.AIcs.CLarXiv:2405.15793v32024
  29. Variational Mixture-of-Experts Autoencoders for Multi-Modal Deep Generative Models

    Yuge Shi, N. Siddharth, Brooks Paige +1

    stat.MLcs.LGarXiv:1911.03393v12019
  30. The Heidelberg spiking datasets for the systematic evaluation of spiking neural networks

    Benjamin Cramer, Yannik Stradmann, Johannes Schemmel +1

    cs.NEcs.LGq-bio.NCarXiv:1910.07407v32019
  31. CAVEWOMAN: How Large Language Models Behave Under Linguistic Input and Output Compression

    Morayo Danielle Adeyemi, Ryan A. Rossi, Franck Dernoncourt

    cs.CLcs.AIcs.LGarXiv:2606.24083v12026
  32. LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding

    Mostafa Elhoushi, Akshat Shrivastava, Diana Liskovich +10

    cs.CLcs.AIcs.LGarXiv:2404.16710v42024
  33. DeepGCNs: Making GCNs Go as Deep as CNNs

    Guohao Li, Matthias Müller, Guocheng Qian +4

    cs.CVcs.LGeess.IVarXiv:1910.06849v32019
  34. The Reward Was in Your Data All Along: Correcting Flow Matching with Discriminator-Guided RL

    Nicolas Beltran-Velez, Felix Friedrich, Zhang Xiaofeng +4

    cs.LGcs.CVarXiv:2606.19162v12026
  35. InceptionTime: Finding AlexNet for Time Series Classification

    Hassan Ismail Fawaz, Benjamin Lucas, Germain Forestier +7

    cs.LGstat.MLarXiv:1909.04939v32019
  36. ControlNet++: Improving Conditional Controls with Efficient Consistency Feedback

    Ming Li, Taojiannan Yang, Huafeng Kuang +4

    cs.CVcs.AIcs.LGarXiv:2404.07987v42024
  37. Anomaly Detection in Video Sequence with Appearance-Motion Correspondence

    Trong Nguyen Nguyen, Jean Meunier

    cs.CVcs.LGcs.NEarXiv:1908.06351v12019
  38. Where Did It Go Wrong? Process-Level Evaluation of Web Agents with Semantic State Tracking

    Jiwan Chung, JiHyuk Byun, Vibhav Vineet +1

    cs.AIcs.LGarXiv:2606.15673v22026
  39. Local Differential Privacy for Deep Learning

    M. A. P. Chamikara, P. Bertok, I. Khalil +3

    cs.LGcs.CRarXiv:1908.02997v32019
  40. Optuna: A Next-generation Hyperparameter Optimization Framework

    Takuya Akiba, Shotaro Sano, Toshihiko Yanase +2

    cs.LGstat.MLarXiv:1907.10902v12019
  41. Attacks, Defenses and Evaluations for LLM Conversation Safety: A Survey

    Zhichen Dong, Zhanhui Zhou, Chao Yang +2

    cs.CLcs.AIcs.CYarXiv:2402.09283v32024
  42. Hierarchically Structured Meta-learning

    Huaxiu Yao, Ying Wei, Junzhou Huang +1

    cs.LGstat.MLarXiv:1905.05301v22019
  43. SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models

    Dongyang Liu, Renrui Zhang, Longtian Qiu +16

    cs.CVcs.AIcs.CLarXiv:2402.05935v32024
  44. InstructIR: High-Quality Image Restoration Following Human Instructions

    Marcos V. Conde, Gregor Geigle, Radu Timofte

    cs.CVcs.LGeess.IVarXiv:2401.16468v52024
  45. A Principle of Targeted Intervention for Multi-Agent Reinforcement Learning

    Anjie Liu, Jianhong Wang, Samuel Kaski +2

    cs.AIcs.LGcs.MAarXiv:2510.17697v42025
  46. Fair Transfer of Multiple Style Attributes in Text

    Karan Dabas, Nishtha Madan, Vijay Arya +3

    cs.CLcs.AIcs.LGarXiv:2001.06693v12020
  47. Convergence guarantees for RMSProp and ADAM in non-convex optimization and an empirical comparison to Nesterov acceleration

    Soham De, Anirbit Mukherjee, Enayat Ullah

    cs.LGmath.OCstat.MLarXiv:1807.06766v32018
  48. Training Language Models with Language Feedback

    Jérémy Scheurer, Jon Ander Campos, Jun Shern Chan +3

    cs.CLcs.AIcs.LGarXiv:2204.14146v42022
  49. PhiNets: Brain-inspired Non-contrastive Learning Based on Temporal Prediction Hypothesis

    Satoki Ishikawa, Makoto Yamada, Han Bao +1

    cs.LGarXiv:2405.14650v22024
  50. Brain-Inspired Stochastic Joint Embedding Representation Learning

    Makoto Yamada, Kian Ming A. Chai, Ayoub Rhim +3

    cs.CVcs.AIcs.LGarXiv:2505.11129v22025
  51. SkyJEPA: Learning Long-Horizon World Models for Zero-Shot Sim-to-Real Control of Quadrotors

    Pratyaksh Rao, Wancong Zhang, Randall Balestriero +2

    cs.ROcs.LGarXiv:2606.23444v22026
  52. Siamese Masked Autoencoders

    Agrim Gupta, Jiajun Wu, Jia Deng +1

    cs.CVcs.LGarXiv:2305.14344v12023
  53. Visual Representation Learning with Stochastic Frame Prediction

    Huiwon Jang, Dongyoung Kim, Junsu Kim +3

    cs.CVcs.AIcs.LGarXiv:2406.07398v22024
  54. Pushing Stochastic Gradient towards Second-Order Methods -- Backpropagation Learning with Transformations in Nonlinearities

    Tommi Vatanen, Tapani Raiko, Harri Valpola +1

    cs.LGcs.CVstat.MLarXiv:1301.3476v32013
  55. TransferTraj: A Vehicle Trajectory Learning Model for Region and Task Transferability

    Tonglong Wei, Yan Lin, Zeyu Zhou +6

    cs.LGarXiv:2505.12672v12025
  56. Stable Learning Using Spiking Neural Networks Equipped With Affine Encoders and Decoders

    A. Martina Neuman, Dominik Dold, Philipp Christian Petersen

    cs.NEcs.LGmath.FAarXiv:2404.04549v32024
  57. Towards Automatic Concept-based Explanations

    Amirata Ghorbani, James Wexler, James Zou +1

    stat.MLcs.CVcs.LGarXiv:1902.03129v32019
  58. Revisiting minimum description length complexity in overparameterized models

    Raaz Dwivedi, Chandan Singh, Bin Yu +1

    cs.LGcs.ITmath.STarXiv:2006.10189v42020
  59. DoMINO: A Decomposable Multi-scale Iterative Neural Operator for Modeling Large Scale Engineering Simulations

    Rishikesh Ranade, Mohammad Amin Nabian, Kaustubh Tangsali +4

    cs.LGphysics.comp-pharXiv:2501.13350v12025
  60. TripNet: Learning Large-scale High-fidelity 3D Car Aerodynamics with Triplane Networks

    Qian Chen, Mohamed Elrefaie, Angela Dai +1

    physics.flu-dyncs.LGarXiv:2503.17400v22025