Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
5,821 to 5,880 of 20,193
Reinforcement Learning and Rule-Based Peer-to-Peer Pricing in Residential PV-BES Communities
Pablo Benalcazar, Maciej Kalka, Wilian Guamán +1
cs.LGcs.CYarXiv:2609.01680v12026Linked Component Analysis from Matrices to High Order Tensors: Applications to Biomedical Data
Guoxu Zhou, Qibin Zhao, Yu Zhang +3
cs.CEcs.LGmath.NAarXiv:1508.07416v12015Learning Human Identity from Motion Patterns
Natalia Neverova, Christian Wolf, Griffin Lacey +4
cs.LGcs.CVcs.NEarXiv:1511.03908v42015AI Research Agents for Machine Learning: Search, Exploration, and Generalization in MLE-bench
Edan Toledo, Karen Hambardzumyan, Martin Josifoski +22
cs.AIcs.LGarXiv:2507.02554v22025Decentralized Federated Learning through Proxy Model Sharing
Shivam Kalra, Junfeng Wen, Jesse C. Cresswell +2
cs.LGarXiv:2111.11343v22021TAPIP3D: Tracking Any Point in Persistent 3D Geometry
Bowei Zhang, Lei Ke, Adam W. Harley +1
cs.CVcs.LGarXiv:2504.14717v32025TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?
Michael S. Ryoo, AJ Piergiovanni, Anurag Arnab +2
cs.CVcs.LGarXiv:2106.11297v42021Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
Ekaterina Fadeeva, Aleksandr Rubashevskii, Artem Shelmanov +9
cs.CLcs.AIcs.LGarXiv:2403.04696v22024EvoX: Meta-Evolution for Automated Discovery
Shu Liu, Shubham Agarwal, Monishwaran Maheswaran +14
cs.LGcs.CLcs.NEarXiv:2602.23413v22026Channel-Wise Attention-Based Network for Self-Supervised Monocular Depth Estimation
Jiaxing Yan, Hong Zhao, Penghui Bu +1
cs.CVcs.AIcs.LGarXiv:2112.13047v12021GenDICE: Generalized Offline Estimation of Stationary Values
Ruiyi Zhang, Bo Dai, Lihong Li +1
stat.MLcs.LGarXiv:2002.09072v12020Design and Analysis of Uplink and Downlink Communications for Federated Learning
Sihui Zheng, Cong Shen, Xiang Chen
cs.ITcs.LGeess.SParXiv:2012.04057v12020SD-LoRA: Scalable Decoupled Low-Rank Adaptation for Class Incremental Learning
Yichen Wu, Hongming Piao, Long-Kai Huang +6
cs.LGarXiv:2501.13198v32025Diffprivlib: The IBM Differential Privacy Library
Naoise Holohan, Stefano Braghin, Pól Mac Aonghusa +1
cs.CRcs.LGarXiv:1907.02444v12019BranchGRPO: Stable and Efficient GRPO with Structured Branching in Diffusion Models
Yuming Li, Yikai Wang, Yuying Zhu +4
cs.CVcs.AIcs.LGarXiv:2509.06040v52025Code-Aware Prompting: A study of Coverage Guided Test Generation in Regression Setting using LLM
Gabriel Ryan, Siddhartha Jain, Mingyue Shang +4
cs.SEcs.LGarXiv:2402.00097v22024Improving the Diffusability of Autoencoders
Ivan Skorokhodov, Sharath Girish, Benran Hu +5
cs.CVcs.AIcs.LGarXiv:2502.14831v32025Preventing Posterior Collapse with delta-VAEs
Ali Razavi, Aäron van den Oord, Ben Poole +1
cs.LGstat.MLarXiv:1901.03416v12019Adjoint Sampling: Highly Scalable Diffusion Samplers via Adjoint Matching
Aaron Havens, Benjamin Kurt Miller, Bing Yan +10
cs.LGcs.AIarXiv:2504.11713v32025Edge-Cloud Collaborative Computing on Distributed Intelligence and Model Optimization: A Survey
Jing Liu, Yao Du, Kun Yang +8
cs.DCcs.AIcs.LGarXiv:2505.01821v52025Imitation Learning as $f$-Divergence Minimization
Liyiming Ke, Sanjiban Choudhury, Matt Barnes +3
cs.LGcs.ITcs.ROarXiv:1905.12888v22019Looking Beyond the Scale: Do Surgical Skill Models Learn Transferable Representations Across Assessment Rubrics?
Hanna Hoffmann, Felix von Bechtolsheim, Stefanie Speidel +1
cs.CVcs.LGarXiv:2608.17519v12026Body size predicts how long ant workers live - but not how they age or how they die from heat
Alana Moscardi, Rafael da Silva, Gleycon Silva
q-bio.PEcs.LGarXiv:2608.14245v12026From Entropy to Epiplexity: Rethinking Information for Computationally Bounded Intelligence
Marc Finzi, Shikai Qiu, Yiding Jiang +3
cs.LGstat.MLarXiv:2601.03220v22026Don't be lazy: CompleteP enables compute-efficient deep transformers
Nolan Dey, Bin Claire Zhang, Lorenzo Noci +6
cs.LGcs.AIarXiv:2505.01618v42025Text Capability Loss in Vision-Language Adaptation: An Attention-Sink Diagnosis
Minsik Choi, Geewook Kim, Young Geun Kim
cs.LGarXiv:2609.00746v12026Web Price Extraction: State of the Art and an Adaptive Browserless Implementation
Evgeniia Kositsyna, Jorge Lloret-Gazo
cs.IRcs.LGcs.NEarXiv:2609.01030v12026A Bayesian Sampling Approach to Exploration in Reinforcement Learning
John Asmuth, Lihong Li, Michael L. Littman +2
cs.LGarXiv:1205.2664v12012Learning Optimal and Fair Decision Trees for Non-Discriminative Decision-Making
Sina Aghaei, Mohammad Javad Azizi, Phebe Vayanos
cs.LGstat.MLarXiv:1903.10598v12019VLA-Cache: Efficient Vision-Language-Action Manipulation via Adaptive Token Caching
Siyu Xu, Yunke Wang, Chenghao Xia +3
cs.ROcs.CVcs.LGarXiv:2502.02175v22025Energy Considerations of Large Language Model Inference and Efficiency Optimizations
Jared Fernandez, Clara Na, Vashisth Tiwari +3
cs.CLcs.LGarXiv:2504.17674v12025Adapting Auxiliary Losses Using Gradient Similarity
Yunshu Du, Wojciech M. Czarnecki, Siddhant M. Jayakumar +3
stat.MLcs.LGarXiv:1812.02224v22018SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning
Rui Pan, Yinwei Dai, Zhihao Zhang +3
cs.LGcs.AIarXiv:2504.07891v22025DMCP: Differentiable Markov Channel Pruning for Neural Networks
Shaopeng Guo, Yujie Wang, Quanquan Li +1
cs.CVcs.LGarXiv:2005.03354v22020Scaling Video Analytics on Constrained Edge Nodes
Christopher Canel, Thomas Kim, Giulio Zhou +5
cs.CVcs.LGcs.PFarXiv:1905.13536v12019Energy-Based Learning for Scene Graph Generation
Mohammed Suhail, Abhay Mittal, Behjat Siddiquie +4
cs.CVcs.LGarXiv:2103.02221v12021SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience
Zeyi Sun, Ziyu Liu, Yuhang Zang +5
cs.AIcs.CLcs.CVarXiv:2508.04700v22025Matched Queries for Curvature and Density at Branching Junctions
Ziqi Zhao, Qingjian Ni
stat.MLcs.LGarXiv:2609.01319v12026Tunable Efficient Unitary Neural Networks (EUNN) and their application to RNNs
Li Jing, Yichen Shen, Tena Dubček +5
cs.LGcs.NEstat.MLarXiv:1612.05231v32016OverThink: Slowdown Attacks on Reasoning LLMs
Abhinav Kumar, Jaechul Roh, Ali Naseh +4
cs.LGcs.CRarXiv:2502.02542v42025AdaMuon: Adaptive Muon Optimizer
Chongjie Si, Debing Zhang, Wei Shen
cs.LGarXiv:2507.11005v32025Natural Neural Networks
Guillaume Desjardins, Karen Simonyan, Razvan Pascanu +1
stat.MLcs.LGcs.NEarXiv:1507.00210v12015Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness
Felix Friedrich, Manuel Brack, Lukas Struppek +4
cs.LGcs.AIcs.CVarXiv:2302.10893v32023Quantum Sparse Autoencoders for Q-Matrix Estimation in Cognitive Diagnosis
Arif Hassan Zidan, Yi Pan, Bowen Guo +5
cs.LGarXiv:2609.01537v12026HarmoCore: Functional Latent Diffusion for Sparse Reconstruction of Oscillatory Wave Fields
Lihao Chen, Xinyu Zhang, Panqi Chen +4
cs.LGcs.CEarXiv:2609.00679v12026MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs
Erik Daxberger, Nina Wenzel, David Griffiths +8
cs.CVcs.CLcs.LGarXiv:2503.13111v22025Radial Attention: $O(n\log n)$ Sparse Attention with Energy Decay for Long Video Generation
Xingyang Li, Muyang Li, Tianle Cai +11
cs.CVcs.AIcs.LGarXiv:2506.19852v22025Rethinking Learnability in Offline Data-driven Optimization
Chao Qian, Chen-Guang Wang, Rong-Xi Tan +1
cs.LGcs.AIcs.NEarXiv:2609.01493v22026What's Behind PPO's Collapse in Long-CoT? Value Optimization Holds the Secret
Yufeng Yuan, Yu Yue, Ruofei Zhu +2
cs.LGarXiv:2503.01491v12025Estimation from Pairwise Comparisons: Sharp Minimax Bounds with Topology Dependence
Nihar B. Shah, Sivaraman Balakrishnan, Joseph Bradley +3
cs.LGcs.ITstat.MLarXiv:1505.01462v12015Fully Parameterized Quantile Function for Distributional Reinforcement Learning
Derek Yang, Li Zhao, Zichuan Lin +3
cs.LGcs.AIstat.MLarXiv:1911.02140v32019A survey of algorithmic recourse: definitions, formulations, solutions, and prospects
Amir-Hossein Karimi, Gilles Barthe, Bernhard Schölkopf +1
cs.LGcs.AIstat.MLarXiv:2010.04050v22020ReasonIR: Training Retrievers for Reasoning Tasks
Rulin Shao, Rui Qiao, Varsha Kishore +8
cs.AIcs.CLcs.IRarXiv:2504.20595v12025TxGemma: Efficient and Agentic LLMs for Therapeutics
Eric Wang, Samuel Schmidgall, Paul F. Jaeger +6
cs.AIcs.CLcs.LGarXiv:2504.06196v12025Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
Hongjin Su, Ruoxi Sun, Jinsung Yoon +3
cs.LGcs.AIarXiv:2501.10893v12025Gluon: Making Muon & Scion Great Again! (Bridging Theory and Practice of LMO-based Optimizers for LLMs)
Artem Riabinin, Egor Shulgin, Kaja Gruntkowska +1
cs.LGmath.OCstat.MLarXiv:2505.13416v12025Adversarial Laser Beam: Effective Physical-World Attack to DNNs in a Blink
Ranjie Duan, Xiaofeng Mao, A. K. Qin +4
cs.LGcs.AIcs.CRarXiv:2103.06504v12021MINE: Towards Continuous Depth MPI with NeRF for Novel View Synthesis
Jiaxin Li, Zijian Feng, Qi She +3
cs.CVcs.GRcs.LGarXiv:2103.14910v32021Self-Training Elicits Concise Reasoning in Large Language Models
Tergel Munkhbat, Namgyu Ho, Seo Hyun Kim +3
cs.CLcs.AIcs.LGarXiv:2502.20122v32025Voice Separation with an Unknown Number of Multiple Speakers
Eliya Nachmani, Yossi Adi, Lior Wolf
eess.AScs.LGcs.SDarXiv:2003.01531v42020