Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,981 to 17,040 of 20,199

  1. A Physical Response-and-Memory Model for Muon Optimization

    Yinze Hu, Hongjun Xiang, Xingao Gong +1

    cs.LGcond-mat.dis-nncond-mat.stat-mecharXiv:2608.22994v12026
  2. MOReL : Model-Based Offline Reinforcement Learning

    Rahul Kidambi, Aravind Rajeswaran, Praneeth Netrapalli +1

    cs.LGcs.AIstat.MLarXiv:2005.05951v32020
  3. Revisiting Spatial-Temporal Similarity: A Deep Learning Framework for Traffic Prediction

    Huaxiu Yao, Xianfeng Tang, Hua Wei +2

    cs.LGarXiv:1803.01254v22018
  4. Knowing Isn't Understanding: Re-grounding Generative Proactivity with Epistemic and Behavioral Insight

    Kirandeep Kaur, Xingda Lyu, Chirag Shah

    cs.CYcs.AIcs.LGarXiv:2602.15259v22026
  5. Multi-Scale Progressive Fusion Network for Single Image Deraining

    Kui Jiang, Zhongyuan Wang, Peng Yi +5

    cs.CVcs.LGeess.IVarXiv:2003.10985v22020
  6. Few-Shot Learning via Embedding Adaptation with Set-to-Set Functions

    Han-Jia Ye, Hexiang Hu, De-Chuan Zhan +1

    cs.LGcs.CVarXiv:1812.03664v62018
  7. Neural Operator based Multi-Field Reconstruction of Inner Solar Boundary State

    Vignesh Kumar Pandian Sathia, Reza Mansouri, Dustin J. Kempton +2

    cs.LGastro-ph.IMastro-ph.SRarXiv:2608.22782v12026
  8. CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

    Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau +4

    cs.CLcs.IRcs.LGarXiv:1911.00359v22019
  9. MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling

    MiniCPM Team, Wenhao An, Yingfa Chen +44

    cs.CLcs.AIcs.LGarXiv:2602.11761v22026
  10. Using Pre-Training Can Improve Model Robustness and Uncertainty

    Dan Hendrycks, Kimin Lee, Mantas Mazeika

    cs.LGcs.CVstat.MLarXiv:1901.09960v52019
  11. InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem

    Shuofei Qiao, Yunxiang Wei, Xuehai Wang +10

    cs.CLcs.AIcs.IRarXiv:2602.14367v22026
  12. Physics-informed neural networks with hard constraints for inverse design

    Lu Lu, Raphael Pestourie, Wenjie Yao +3

    physics.comp-phcs.LGarXiv:2102.04626v12021
  13. Attack of the Tails: Yes, You Really Can Backdoor Federated Learning

    Hongyi Wang, Kartik Sreenivasan, Shashank Rajput +5

    cs.LGcs.CRcs.DCarXiv:2007.05084v12020
  14. DF-MoE: Generalizable Deepfake Detection via Multimodal Sparse Mixture-of-Experts

    Vlad Hondru, Florinel Alin Croitoru, Iuliana Georgescu +2

    cs.CVcs.AIcs.LGarXiv:2608.23363v12026
  15. LoST: Level of Semantics Tokenization for 3D Shapes

    Niladri Shekhar Dutt, Zifan Shi, Paul Guerrero +4

    cs.CVcs.GRcs.LGarXiv:2603.17995v12026
  16. HopSkipJumpAttack: A Query-Efficient Decision-Based Attack

    Jianbo Chen, Michael I. Jordan, Martin J. Wainwright

    cs.LGcs.CRmath.OCarXiv:1904.02144v52019
  17. Reservoir of Importance: Learning Semi-Structured Sparsity with Differentiable Subset Sampling

    Ha Dinh, Xuan Duy Ta, Khoat Than +1

    cs.LGarXiv:2608.23048v12026
  18. Spanning the Visual Analogy Space with a Weight Basis of LoRAs

    Hila Manor, Rinon Gal, Haggai Maron +2

    cs.CVcs.AIcs.GRarXiv:2602.15727v22026
  19. CMI-RewardBench: Evaluating Music Reward Models with Compositional Multimodal Instruction

    Yinghao Ma, Haiwen Xia, Hewei Gao +9

    cs.SDcs.AIcs.LGarXiv:2603.00610v32026
  20. Segment Anything Model for Medical Image Analysis: an Experimental Study

    Maciej A. Mazurowski, Haoyu Dong, Hanxue Gu +3

    cs.CVcs.AIcs.LGarXiv:2304.10517v32023
  21. Contrastive Clustering

    Yunfan Li, Peng Hu, Zitao Liu +3

    cs.LGcs.CVstat.MLarXiv:2009.09687v12020
  22. Spend Search Where It Pays: Value-Guided Structured Sampling and Optimization for Generative Recommendation

    Jie Jiang, Yangru Huang, Zeyu Wang +4

    cs.AIcs.LGarXiv:2602.10699v22026
  23. KILT: a Benchmark for Knowledge Intensive Language Tasks

    Fabio Petroni, Aleksandra Piktus, Angela Fan +10

    cs.CLcs.AIcs.IRarXiv:2009.02252v42020
  24. High-Fidelity Audio Compression with Improved RVQGAN

    Rithesh Kumar, Prem Seetharaman, Alejandro Luebs +2

    cs.SDcs.LGeess.ASarXiv:2306.06546v22023
  25. Transfer learning enhanced physics informed neural network for phase-field modeling of fracture

    Somdatta Goswami, Cosmin Anitescu, Souvik Chakraborty +1

    stat.MLcs.LGarXiv:1907.02531v12019
  26. An introduction to Topological Data Analysis: fundamental and practical aspects for data scientists

    Frédéric Chazal, Bertrand Michel

    math.STcs.LGmath.ATarXiv:1710.04019v22017
  27. Online Learning: A Comprehensive Survey

    Steven C. H. Hoi, Doyen Sahoo, Jing Lu +1

    cs.LGarXiv:1802.02871v22018
  28. Counterfactual Transition Graphs: Evaluating Cross-Class Transition Quality

    Syed Muhammad Hamza Zaidi, Szymon Bobek, Grzegorz J. Nalepa +1

    cs.LGcs.AIarXiv:2608.23164v12026
  29. PixelDefend: Leveraging Generative Models to Understand and Defend against Adversarial Examples

    Yang Song, Taesup Kim, Sebastian Nowozin +2

    cs.LGarXiv:1710.10766v32017
  30. Mirror descent algorithms with logarithmic barriers

    Alberto De Marchi, Yura Malitsky, Adrien B. Taylor

    math.OCcs.LGmath.NAarXiv:2608.22834v12026
  31. Transformers learn in-context by gradient descent

    Johannes von Oswald, Eyvind Niklasson, Ettore Randazzo +4

    cs.LGcs.AIcs.CLarXiv:2212.07677v22022
  32. Model compression via distillation and quantization

    Antonio Polino, Razvan Pascanu, Dan Alistarh

    cs.NEcs.LGarXiv:1802.05668v12018
  33. Likelihood Ratios for Out-of-Distribution Detection

    Jie Ren, Peter J. Liu, Emily Fertig +5

    stat.MLcs.LGarXiv:1906.02845v22019
  34. Next Embedding Prediction Makes World Models Stronger

    George Bredis, Nikita Balagansky, Daniil Gavrilov +1

    cs.LGcs.AIarXiv:2603.02765v12026
  35. Reward-Free Continual Adaptation for Resilient Space Robots

    Andrej Orsula, Miguel Olivares-Mendez, Carol Martinez

    cs.ROcs.AIcs.LGarXiv:2608.23452v12026
  36. Machine learning for molecular simulation

    Frank Noé, Alexandre Tkatchenko, Klaus-Robert Müller +1

    physics.chem-phcs.LGphysics.comp-pharXiv:1911.02792v12019
  37. ISO-Bench: Can Coding Agents Optimize Real-World Inference Workloads?

    Ayush Nangia, Shikhar Mishra, Aman Gokrani +1

    cs.LGarXiv:2602.19594v12026
  38. BeamPERL: Parameter-Efficient RL with Verifiable Rewards Specializes Compact LLMs for Structured Beam Mechanics Reasoning

    Tarjei Paule Hage, Markus J. Buehler

    cs.AIcond-mat.mtrl-scics.CLarXiv:2603.04124v12026
  39. Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

    Jannis Becktepe, Aleksandra Franz, Nils Thuerey +1

    cs.LGarXiv:2601.15015v22026
  40. Document Ranking with a Pretrained Sequence-to-Sequence Model

    Rodrigo Nogueira, Zhiying Jiang, Jimmy Lin

    cs.IRcs.LGarXiv:2003.06713v12020
  41. Diffusion Model Alignment Using Direct Preference Optimization

    Bram Wallace, Meihua Dang, Rafael Rafailov +7

    cs.CVcs.AIcs.GRarXiv:2311.12908v12023
  42. DenseCLIP: Language-Guided Dense Prediction with Context-Aware Prompting

    Yongming Rao, Wenliang Zhao, Guangyi Chen +5

    cs.CVcs.AIcs.LGarXiv:2112.01518v22021
  43. What Can Transformers Learn In-Context? A Case Study of Simple Function Classes

    Shivam Garg, Dimitris Tsipras, Percy Liang +1

    cs.CLcs.LGarXiv:2208.01066v32022
  44. Bridging Theory and Algorithm for Domain Adaptation

    Yuchen Zhang, Tianle Liu, Mingsheng Long +1

    cs.LGstat.MLarXiv:1904.05801v22019
  45. Don't Repeat Yourself: Stopping Verbatim Loops at Sampling Time

    Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder +2

    cs.CLcs.AIcs.LGarXiv:2608.22761v12026
  46. A Commutator Framework for Selective Spectral Alignment in Deep Neural Networks

    Kaj Nyström

    stat.MLcs.LGarXiv:2608.22910v12026
  47. Parameterized Explainer for Graph Neural Network

    Dongsheng Luo, Wei Cheng, Dongkuan Xu +4

    cs.LGcs.AIarXiv:2011.04573v12020
  48. Beyond chlorophyll: machine learning estimates of diagnostic phytoplankton pigments from multispectral ocean colour data

    David Moffat, Angus Laurenson, Victor Martinez-Vicente +4

    q-bio.OTcs.LGphysics.opticsarXiv:2608.23348v12026
  49. Thinking at the Right Size: Amortized Distillation Across Post-Trained LLMs

    Yan Zhou, Sara Kangaslahti, Jonathan Geuter +4

    cs.LGarXiv:2608.22854v12026
  50. Least-Loaded Expert Parallelism: Load Balancing An Imbalanced Mixture-of-Experts

    Xuan-Phi Nguyen, Shrey Pandit, Austin Xu +2

    cs.LGcs.AIarXiv:2601.17111v12026
  51. Conditional Positional Encodings for Vision Transformers

    Xiangxiang Chu, Zhi Tian, Bo Zhang +2

    cs.CVcs.AIcs.LGarXiv:2102.10882v32021
  52. VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning?

    Minkyu Kim, Sangheon Lee, Dongmin Park

    cs.CVcs.AIcs.LGarXiv:2603.07888v12026
  53. Meta-Learning: A Survey

    Joaquin Vanschoren

    cs.LGstat.MLarXiv:1810.03548v12018
  54. XTC: Head-Aware Sampling by Excluding Top Choices

    Philipp Emanuel Weidmann, Allen Roush, Judah Goldfeder +2

    cs.CLcs.AIcs.LGarXiv:2608.22758v12026
  55. On Distillation of Guided Diffusion Models

    Chenlin Meng, Robin Rombach, Ruiqi Gao +4

    cs.CVcs.AIcs.LGarXiv:2210.03142v32022
  56. TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning

    Alliot Nagle, Jakhongir Saydaliev, Dhia Garbaya +3

    cs.LGcs.AIcs.CLarXiv:2603.12529v22026
  57. The Web as a Knowledge-base for Answering Complex Questions

    Alon Talmor, Jonathan Berant

    cs.CLcs.AIcs.LGarXiv:1803.06643v12018
  58. PolyChirp: Multi-Species Birdsong Classification Using TinyML on Low-Power Acoustic Sensors

    Nathan Duboisset, Zhaolan Huang, Felix Bießmann +3

    cs.LGcs.AIarXiv:2608.23101v12026
  59. Algorithms for nonnegative matrix factorization with the beta-divergence

    Cédric Févotte, Jérôme Idier

    cs.LGarXiv:1010.1763v32010
  60. COVID-19 Image Data Collection: Prospective Predictions Are the Future

    Joseph Paul Cohen, Paul Morrison, Lan Dao +3

    q-bio.QMcs.CVcs.LGarXiv:2006.11988v32020