Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

5,821 to 5,880 of 20,193

  1. Reinforcement Learning and Rule-Based Peer-to-Peer Pricing in Residential PV-BES Communities

    Pablo Benalcazar, Maciej Kalka, Wilian Guamán +1

    cs.LGcs.CYarXiv:2609.01680v12026
  2. Linked Component Analysis from Matrices to High Order Tensors: Applications to Biomedical Data

    Guoxu Zhou, Qibin Zhao, Yu Zhang +3

    cs.CEcs.LGmath.NAarXiv:1508.07416v12015
  3. Learning Human Identity from Motion Patterns

    Natalia Neverova, Christian Wolf, Griffin Lacey +4

    cs.LGcs.CVcs.NEarXiv:1511.03908v42015
  4. AI Research Agents for Machine Learning: Search, Exploration, and Generalization in MLE-bench

    Edan Toledo, Karen Hambardzumyan, Martin Josifoski +22

    cs.AIcs.LGarXiv:2507.02554v22025
  5. Decentralized Federated Learning through Proxy Model Sharing

    Shivam Kalra, Junfeng Wen, Jesse C. Cresswell +2

    cs.LGarXiv:2111.11343v22021
  6. TAPIP3D: Tracking Any Point in Persistent 3D Geometry

    Bowei Zhang, Lei Ke, Adam W. Harley +1

    cs.CVcs.LGarXiv:2504.14717v32025
  7. TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?

    Michael S. Ryoo, AJ Piergiovanni, Anurag Arnab +2

    cs.CVcs.LGarXiv:2106.11297v42021
  8. Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification

    Ekaterina Fadeeva, Aleksandr Rubashevskii, Artem Shelmanov +9

    cs.CLcs.AIcs.LGarXiv:2403.04696v22024
  9. EvoX: Meta-Evolution for Automated Discovery

    Shu Liu, Shubham Agarwal, Monishwaran Maheswaran +14

    cs.LGcs.CLcs.NEarXiv:2602.23413v22026
  10. Channel-Wise Attention-Based Network for Self-Supervised Monocular Depth Estimation

    Jiaxing Yan, Hong Zhao, Penghui Bu +1

    cs.CVcs.AIcs.LGarXiv:2112.13047v12021
  11. GenDICE: Generalized Offline Estimation of Stationary Values

    Ruiyi Zhang, Bo Dai, Lihong Li +1

    stat.MLcs.LGarXiv:2002.09072v12020
  12. Design and Analysis of Uplink and Downlink Communications for Federated Learning

    Sihui Zheng, Cong Shen, Xiang Chen

    cs.ITcs.LGeess.SParXiv:2012.04057v12020
  13. SD-LoRA: Scalable Decoupled Low-Rank Adaptation for Class Incremental Learning

    Yichen Wu, Hongming Piao, Long-Kai Huang +6

    cs.LGarXiv:2501.13198v32025
  14. Diffprivlib: The IBM Differential Privacy Library

    Naoise Holohan, Stefano Braghin, Pól Mac Aonghusa +1

    cs.CRcs.LGarXiv:1907.02444v12019
  15. BranchGRPO: Stable and Efficient GRPO with Structured Branching in Diffusion Models

    Yuming Li, Yikai Wang, Yuying Zhu +4

    cs.CVcs.AIcs.LGarXiv:2509.06040v52025
  16. Code-Aware Prompting: A study of Coverage Guided Test Generation in Regression Setting using LLM

    Gabriel Ryan, Siddhartha Jain, Mingyue Shang +4

    cs.SEcs.LGarXiv:2402.00097v22024
  17. Improving the Diffusability of Autoencoders

    Ivan Skorokhodov, Sharath Girish, Benran Hu +5

    cs.CVcs.AIcs.LGarXiv:2502.14831v32025
  18. Preventing Posterior Collapse with delta-VAEs

    Ali Razavi, Aäron van den Oord, Ben Poole +1

    cs.LGstat.MLarXiv:1901.03416v12019
  19. Adjoint Sampling: Highly Scalable Diffusion Samplers via Adjoint Matching

    Aaron Havens, Benjamin Kurt Miller, Bing Yan +10

    cs.LGcs.AIarXiv:2504.11713v32025
  20. Edge-Cloud Collaborative Computing on Distributed Intelligence and Model Optimization: A Survey

    Jing Liu, Yao Du, Kun Yang +8

    cs.DCcs.AIcs.LGarXiv:2505.01821v52025
  21. Imitation Learning as $f$-Divergence Minimization

    Liyiming Ke, Sanjiban Choudhury, Matt Barnes +3

    cs.LGcs.ITcs.ROarXiv:1905.12888v22019
  22. Looking Beyond the Scale: Do Surgical Skill Models Learn Transferable Representations Across Assessment Rubrics?

    Hanna Hoffmann, Felix von Bechtolsheim, Stefanie Speidel +1

    cs.CVcs.LGarXiv:2608.17519v12026
  23. Body size predicts how long ant workers live - but not how they age or how they die from heat

    Alana Moscardi, Rafael da Silva, Gleycon Silva

    q-bio.PEcs.LGarXiv:2608.14245v12026
  24. From Entropy to Epiplexity: Rethinking Information for Computationally Bounded Intelligence

    Marc Finzi, Shikai Qiu, Yiding Jiang +3

    cs.LGstat.MLarXiv:2601.03220v22026
  25. Don't be lazy: CompleteP enables compute-efficient deep transformers

    Nolan Dey, Bin Claire Zhang, Lorenzo Noci +6

    cs.LGcs.AIarXiv:2505.01618v42025
  26. Text Capability Loss in Vision-Language Adaptation: An Attention-Sink Diagnosis

    Minsik Choi, Geewook Kim, Young Geun Kim

    cs.LGarXiv:2609.00746v12026
  27. Web Price Extraction: State of the Art and an Adaptive Browserless Implementation

    Evgeniia Kositsyna, Jorge Lloret-Gazo

    cs.IRcs.LGcs.NEarXiv:2609.01030v12026
  28. A Bayesian Sampling Approach to Exploration in Reinforcement Learning

    John Asmuth, Lihong Li, Michael L. Littman +2

    cs.LGarXiv:1205.2664v12012
  29. Learning Optimal and Fair Decision Trees for Non-Discriminative Decision-Making

    Sina Aghaei, Mohammad Javad Azizi, Phebe Vayanos

    cs.LGstat.MLarXiv:1903.10598v12019
  30. VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching

    Siyu Xu, Yunke Wang, Chenghao Xia +3

    cs.ROcs.CVcs.LGarXiv:2502.02175v22025
  31. Energy Considerations of Large Language Model Inference and Efficiency Optimizations

    Jared Fernandez, Clara Na, Vashisth Tiwari +3

    cs.CLcs.LGarXiv:2504.17674v12025
  32. Adapting Auxiliary Losses Using Gradient Similarity

    Yunshu Du, Wojciech M. Czarnecki, Siddhant M. Jayakumar +3

    stat.MLcs.LGarXiv:1812.02224v22018
  33. SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning

    Rui Pan, Yinwei Dai, Zhihao Zhang +3

    cs.LGcs.AIarXiv:2504.07891v22025
  34. DMCP: Differentiable Markov Channel Pruning for Neural Networks

    Shaopeng Guo, Yujie Wang, Quanquan Li +1

    cs.CVcs.LGarXiv:2005.03354v22020
  35. Scaling Video Analytics on Constrained Edge Nodes

    Christopher Canel, Thomas Kim, Giulio Zhou +5

    cs.CVcs.LGcs.PFarXiv:1905.13536v12019
  36. Energy-Based Learning for Scene Graph Generation

    Mohammed Suhail, Abhay Mittal, Behjat Siddiquie +4

    cs.CVcs.LGarXiv:2103.02221v12021
  37. SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience

    Zeyi Sun, Ziyu Liu, Yuhang Zang +5

    cs.AIcs.CLcs.CVarXiv:2508.04700v22025
  38. Matched Queries for Curvature and Density at Branching Junctions

    Ziqi Zhao, Qingjian Ni

    stat.MLcs.LGarXiv:2609.01319v12026
  39. Tunable Efficient Unitary Neural Networks (EUNN) and their application to RNNs

    Li Jing, Yichen Shen, Tena Dubček +5

    cs.LGcs.NEstat.MLarXiv:1612.05231v32016
  40. OverThink: Slowdown Attacks on Reasoning LLMs

    Abhinav Kumar, Jaechul Roh, Ali Naseh +4

    cs.LGcs.CRarXiv:2502.02542v42025
  41. AdaMuon: Adaptive Muon Optimizer

    Chongjie Si, Debing Zhang, Wei Shen

    cs.LGarXiv:2507.11005v32025
  42. Natural Neural Networks

    Guillaume Desjardins, Karen Simonyan, Razvan Pascanu +1

    stat.MLcs.LGcs.NEarXiv:1507.00210v12015
  43. Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness

    Felix Friedrich, Manuel Brack, Lukas Struppek +4

    cs.LGcs.AIcs.CVarXiv:2302.10893v32023
  44. Quantum Sparse Autoencoders for Q-Matrix Estimation in Cognitive Diagnosis

    Arif Hassan Zidan, Yi Pan, Bowen Guo +5

    cs.LGarXiv:2609.01537v12026
  45. HarmoCore: Functional Latent Diffusion for Sparse Reconstruction of Oscillatory Wave Fields

    Lihao Chen, Xinyu Zhang, Panqi Chen +4

    cs.LGcs.CEarXiv:2609.00679v12026
  46. MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

    Erik Daxberger, Nina Wenzel, David Griffiths +8

    cs.CVcs.CLcs.LGarXiv:2503.13111v22025
  47. Radial Attention: $O(n\log n)$ Sparse Attention with Energy Decay for Long Video Generation

    Xingyang Li, Muyang Li, Tianle Cai +11

    cs.CVcs.AIcs.LGarXiv:2506.19852v22025
  48. Rethinking Learnability in Offline Data-driven Optimization

    Chao Qian, Chen-Guang Wang, Rong-Xi Tan +1

    cs.LGcs.AIcs.NEarXiv:2609.01493v22026
  49. What's Behind PPO's Collapse in Long-CoT? Value Optimization Holds the Secret

    Yufeng Yuan, Yu Yue, Ruofei Zhu +2

    cs.LGarXiv:2503.01491v12025
  50. Estimation from Pairwise Comparisons: Sharp Minimax Bounds with Topology Dependence

    Nihar B. Shah, Sivaraman Balakrishnan, Joseph Bradley +3

    cs.LGcs.ITstat.MLarXiv:1505.01462v12015
  51. Fully Parameterized Quantile Function for Distributional Reinforcement Learning

    Derek Yang, Li Zhao, Zichuan Lin +3

    cs.LGcs.AIstat.MLarXiv:1911.02140v32019
  52. A survey of algorithmic recourse: definitions, formulations, solutions, and prospects

    Amir-Hossein Karimi, Gilles Barthe, Bernhard Schölkopf +1

    cs.LGcs.AIstat.MLarXiv:2010.04050v22020
  53. ReasonIR: Training Retrievers for Reasoning Tasks

    Rulin Shao, Rui Qiao, Varsha Kishore +8

    cs.AIcs.CLcs.IRarXiv:2504.20595v12025
  54. TxGemma: Efficient and Agentic LLMs for Therapeutics

    Eric Wang, Samuel Schmidgall, Paul F. Jaeger +6

    cs.AIcs.CLcs.LGarXiv:2504.06196v12025
  55. Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments

    Hongjin Su, Ruoxi Sun, Jinsung Yoon +3

    cs.LGcs.AIarXiv:2501.10893v12025
  56. Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)

    Artem Riabinin, Egor Shulgin, Kaja Gruntkowska +1

    cs.LGmath.OCstat.MLarXiv:2505.13416v12025
  57. Adversarial Laser Beam: Effective Physical-World Attack to DNNs in a Blink

    Ranjie Duan, Xiaofeng Mao, A. K. Qin +4

    cs.LGcs.AIcs.CRarXiv:2103.06504v12021
  58. MINE: Towards Continuous Depth MPI with NeRF for Novel View Synthesis

    Jiaxin Li, Zijian Feng, Qi She +3

    cs.CVcs.GRcs.LGarXiv:2103.14910v32021
  59. Self-Training Elicits Concise Reasoning in Large Language Models

    Tergel Munkhbat, Namgyu Ho, Seo Hyun Kim +3

    cs.CLcs.AIcs.LGarXiv:2502.20122v32025
  60. Voice Separation with an Unknown Number of Multiple Speakers

    Eliya Nachmani, Yossi Adi, Lior Wolf

    eess.AScs.LGcs.SDarXiv:2003.01531v42020