Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

16,861 to 16,920 of 20,217

  1. No More Strided Convolutions or Pooling: A New CNN Building Block for Low-Resolution Images and Small Objects

    Raja Sunkara, Tie Luo

    cs.CVcs.LGarXiv:2208.03641v12022
  2. DiffDock: Diffusion Steps, Twists, and Turns for Molecular Docking

    Gabriele Corso, Hannes Stärk, Bowen Jing +2

    q-bio.BMcs.LGphysics.bio-pharXiv:2210.01776v22022
  3. Bridging Academia and Industry: A Comprehensive Benchmark for Attributed Graph Clustering

    Yunhui Liu, Pengyu Qiu, Yu Xing +6

    cs.LGarXiv:2602.08519v12026
  4. Effective Reasoning Chains Reduce Intrinsic Dimensionality

    Archiki Prasad, Mandar Joshi, Kenton Lee +2

    cs.CLcs.AIcs.LGarXiv:2602.09276v22026
  5. From Multimodal Observation to Interpretable Suggestions: Counterfactual Time-Expanded Relational Modeling of Surgical Teams

    Vincenzo Marco De Luca, Antonio Longa, Giovanna Varni +1

    cs.LGarXiv:2608.23254v12026
  6. Object Goal Navigation using Goal-Oriented Semantic Exploration

    Devendra Singh Chaplot, Dhiraj Gandhi, Abhinav Gupta +1

    cs.CVcs.LGcs.ROarXiv:2007.00643v22020
  7. NeST: Neuron Selective Tuning for LLM Safety

    Sasha Behrouzi, Lichao Wu, Mohamadreza Rostami +1

    cs.CRcs.LGarXiv:2602.16835v22026
  8. BridgeData V2: A Dataset for Robot Learning at Scale

    Homer Walke, Kevin Black, Abraham Lee +11

    cs.ROcs.LGarXiv:2308.12952v32023
  9. Reinforced Attention Learning

    Bangzheng Li, Jianmo Ni, Chen Qu +5

    cs.CLcs.CVcs.LGarXiv:2602.04884v22026
  10. Self-Adaptive Physics-Informed Neural Networks using a Soft Attention Mechanism

    Levi McClenny, Ulisses Braga-Neto

    cs.LGstat.MLarXiv:2009.04544v52020
  11. MotionCrafter: Dense Geometry and Motion Reconstruction with a 4D VAE

    Ruijie Zhu, Jiahao Lu, Wenbo Hu +4

    cs.CVcs.AIcs.CGarXiv:2602.08961v22026
  12. Towards Actionable Surgical Team Dynamics: from Teamwork to Counterfactual Annotations

    Vincenzo Marco De Luca, Antonio Longa, Andrea Passerini

    cs.LGarXiv:2608.23344v12026
  13. Joint Optimization Framework for Learning with Noisy Labels

    Daiki Tanaka, Daiki Ikami, Toshihiko Yamasaki +1

    cs.CVcs.LGstat.MLarXiv:1803.11364v12018
  14. SMASH: One-Shot Model Architecture Search through HyperNetworks

    Andrew Brock, Theodore Lim, J. M. Ritchie +1

    cs.LGarXiv:1708.05344v12017
  15. Reasoning-Augmented Representations for Multimodal Retrieval

    Jianrui Zhang, Anirudh Sundara Rajan, Brandon Han +3

    cs.IRcs.AIcs.CVarXiv:2602.07125v12026
  16. Reinforcement Learning in Healthcare: A Survey

    Chao Yu, Jiming Liu, Shamim Nemati

    cs.LGcs.AIarXiv:1908.08796v42019
  17. Spectral Temporal Graph Neural Network for Multivariate Time-series Forecasting

    Defu Cao, Yujing Wang, Juanyong Duan +8

    cs.LGcs.AIarXiv:2103.07719v12021
  18. Stemphonic: All-at-once Flexible Multi-stem Music Generation

    Shih-Lun Wu, Ge Zhu, Juan-Pablo Caceres +2

    cs.SDcs.LGcs.MMarXiv:2602.09891v12026
  19. Hardware Co-Design Scaling Laws via Roofline Modelling for On-Device LLMs

    Luoyang Sun, Jiwen Jiang, Yifeng Ding +9

    cs.LGcs.CLarXiv:2602.10377v12026
  20. Structured Pruning of Deep Convolutional Neural Networks

    Sajid Anwar, Kyuyeon Hwang, Wonyong Sung

    cs.NEcs.LGstat.MLarXiv:1512.08571v12015
  21. A Survey of Inverse Reinforcement Learning: Challenges, Methods and Progress

    Saurabh Arora, Prashant Doshi

    cs.LGstat.MLarXiv:1806.06877v32018
  22. Video Frame Synthesis using Deep Voxel Flow

    Ziwei Liu, Raymond A. Yeh, Xiaoou Tang +2

    cs.CVcs.GRcs.LGarXiv:1702.02463v22017
  23. Multi-Task Learning with Deep Neural Networks: A Survey

    Michael Crawshaw

    cs.LGcs.CVstat.MLarXiv:2009.09796v12020
  24. UniT: Unified Multimodal Chain-of-Thought Test-time Scaling

    Leon Liangyu Chen, Haoyu Ma, Zhipeng Fan +11

    cs.CVcs.AIcs.LGarXiv:2602.12279v22026
  25. The (Un)reliability of saliency methods

    Pieter-Jan Kindermans, Sara Hooker, Julius Adebayo +5

    stat.MLcs.LGarXiv:1711.00867v12017
  26. DeepVision-103K: A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning

    Haoxiang Sun, Lizhen Xu, Bing Zhao +5

    cs.LGcs.AIarXiv:2602.16742v12026
  27. Human Pose Estimation with Iterative Error Feedback

    Joao Carreira, Pulkit Agrawal, Katerina Fragkiadaki +1

    cs.CVcs.LGcs.NEarXiv:1507.06550v32015
  28. Weight Decay Improves Language Model Plasticity

    Tessa Han, Sebastian Bordt, Hanlin Zhang +1

    cs.LGcs.AIcs.CLarXiv:2602.11137v22026
  29. Real-Time Flying Object Detection with YOLOv8

    Dillon Reis, Jordan Kupec, Jacqueline Hong +1

    cs.CVcs.LGarXiv:2305.09972v22023
  30. ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces

    Xin Xu, Tong Yu, Xiang Chen +3

    cs.AIcs.CLcs.LGarXiv:2602.11683v12026
  31. Variational Continual Learning

    Cuong V. Nguyen, Yingzhen Li, Thang D. Bui +1

    stat.MLcs.LGarXiv:1710.10628v32017
  32. Unsupervised Neural Machine Translation

    Mikel Artetxe, Gorka Labaka, Eneko Agirre +1

    cs.CLcs.AIcs.LGarXiv:1710.11041v22017
  33. Vector-quantized Image Modeling with Improved VQGAN

    Jiahui Yu, Xin Li, Jing Yu Koh +7

    cs.CVcs.LGarXiv:2110.04627v32021
  34. MAEB: Massive Audio Embedding Benchmark

    Adnan El Assadi, Isaac Chung, Chenghao Xiao +15

    cs.SDcs.AIcs.CLarXiv:2602.16008v12026
  35. Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

    Yunfei Chu, Jin Xu, Xiaohuan Zhou +5

    eess.AScs.CLcs.LGarXiv:2311.07919v22023
  36. Using millions of emoji occurrences to learn any-domain representations for detecting sentiment, emotion and sarcasm

    Bjarke Felbo, Alan Mislove, Anders Søgaard +2

    stat.MLcs.LGarXiv:1708.00524v22017
  37. Linear Mode Connectivity and the Lottery Ticket Hypothesis

    Jonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy +1

    cs.LGcs.NEstat.MLarXiv:1912.05671v42019
  38. Sink-Aware Pruning for Diffusion Language Models

    Aidar Myrzakhan, Tianyi Li, Bowei Guo +2

    cs.CLcs.AIcs.LGarXiv:2602.17664v12026
  39. Reinforcement Learning for Combinatorial Optimization: A Survey

    Nina Mazyavkina, Sergey Sviridov, Sergei Ivanov +1

    cs.LGmath.COmath.OCarXiv:2003.03600v32020
  40. Synthetic and Natural Noise Both Break Neural Machine Translation

    Yonatan Belinkov, Yonatan Bisk

    cs.CLcs.LGarXiv:1711.02173v22017
  41. Beyond the Harness: End-to-End Optimization of Context Artifacts for Enterprise Text-to-SQL

    Kate Gwimm, Carson Eisenach

    cs.AIcs.LGarXiv:2608.22830v12026
  42. code2seq: Generating Sequences from Structured Representations of Code

    Uri Alon, Shaked Brody, Omer Levy +1

    cs.LGcs.PLstat.MLarXiv:1808.01400v62018
  43. ExpeL: LLM Agents Are Experiential Learners

    Andrew Zhao, Daniel Huang, Quentin Xu +3

    cs.LGcs.AIcs.CLarXiv:2308.10144v32023
  44. On the "Induction Bias" in Sequence Models

    M. Reza Ebrahimi, Michaël Defferrard, Sunny Panchal +1

    cs.LGcs.CLarXiv:2602.18333v22026
  45. Nesterov Accelerated Gradient and Scale Invariance for Adversarial Attacks

    Jiadong Lin, Chuanbiao Song, Kun He +2

    cs.LGcs.CRstat.MLarXiv:1908.06281v52019
  46. LFPO: Likelihood-Free Policy Optimization for Masked Diffusion Models

    Chenxing Wei, Jiazhen Kang, Hong Wang +8

    cs.LGcs.AIarXiv:2603.01563v12026
  47. Transition-Based Dependency Parsing with Stack Long Short-Term Memory

    Chris Dyer, Miguel Ballesteros, Wang Ling +2

    cs.CLcs.LGcs.NEarXiv:1505.08075v12015
  48. MMD GAN: Towards Deeper Understanding of Moment Matching Network

    Chun-Liang Li, Wei-Cheng Chang, Yu Cheng +2

    cs.LGcs.AIstat.MLarXiv:1705.08584v32017
  49. Test-Time Adaptation for ECG Classification via SQI-Gated Self-Training and Beat-Rhythm Consistency

    Wenhan Jiang, Zhipeng Deng, Jiale Zhou +3

    cs.LGarXiv:2608.23347v12026
  50. Virtual-to-real Deep Reinforcement Learning: Continuous Control of Mobile Robots for Mapless Navigation

    Lei Tai, Giuseppe Paolo, Ming Liu

    cs.ROcs.AIcs.LGarXiv:1703.00420v42017
  51. SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching

    Yasaman Haghighi, Alexandre Alahi

    cs.CVcs.LGarXiv:2602.24208v12026
  52. PTE: Predictive Text Embedding through Large-scale Heterogeneous Text Networks

    Jian Tang, Meng Qu, Qiaozhu Mei

    cs.CLcs.LGcs.NEarXiv:1508.00200v12015
  53. A Brief Review of Domain Adaptation

    Abolfazl Farahani, Sahar Voghoei, Khaled Rasheed +1

    cs.LGcs.CVarXiv:2010.03978v12020
  54. Agentic Critical Training

    Weize Liu, Minghui Liu, Sy-Tuyen Ho +3

    cs.AIcs.CLcs.LGarXiv:2603.08706v12026
  55. Consistency of the group Lasso and multiple kernel learning

    Francis Bach

    cs.LGarXiv:0707.3390v22007
  56. MetaCaster: Meta-Harness-Optimized Agent for End-to-End Few-Shot Learning of Lightweight Time Series Forecasters

    ChengAo Shen, Wenchao Yu, Fangyu Wu +6

    cs.LGcs.AIarXiv:2608.23473v12026
  57. Human-level performance in first-person multiplayer games with population-based deep reinforcement learning

    Max Jaderberg, Wojciech M. Czarnecki, Iain Dunning +15

    cs.LGcs.AIstat.MLarXiv:1807.01281v12018
  58. How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities

    Ziwen Xu, Kewei Xu, Haoming Xu +8

    cs.CLcs.AIcs.HCarXiv:2603.02578v22026
  59. Random Search and Reproducibility for Neural Architecture Search

    Liam Li, Ameet Talwalkar

    cs.LGstat.MLarXiv:1902.07638v32019
  60. TourPlanner: A Competitive Consensus Framework with Constraint-Gated Reinforcement Learning for Travel Planning

    Yinuo Wang, Mining Tan, Wenxiang Jiao +5

    cs.AIcs.CLcs.LGarXiv:2601.04698v12026