Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
16,861 to 16,920 of 20,217
No More Strided Convolutions or Pooling: A New CNN Building Block for Low-Resolution Images and Small Objects
Raja Sunkara, Tie Luo
cs.CVcs.LGarXiv:2208.03641v12022DiffDock: Diffusion Steps, Twists, and Turns for Molecular Docking
Gabriele Corso, Hannes Stärk, Bowen Jing +2
q-bio.BMcs.LGphysics.bio-pharXiv:2210.01776v22022Bridging Academia and Industry: A Comprehensive Benchmark for Attributed Graph Clustering
Yunhui Liu, Pengyu Qiu, Yu Xing +6
cs.LGarXiv:2602.08519v12026Effective Reasoning Chains Reduce Intrinsic Dimensionality
Archiki Prasad, Mandar Joshi, Kenton Lee +2
cs.CLcs.AIcs.LGarXiv:2602.09276v22026From Multimodal Observation to Interpretable Suggestions: Counterfactual Time-Expanded Relational Modeling of Surgical Teams
Vincenzo Marco De Luca, Antonio Longa, Giovanna Varni +1
cs.LGarXiv:2608.23254v12026Object Goal Navigation using Goal-Oriented Semantic Exploration
Devendra Singh Chaplot, Dhiraj Gandhi, Abhinav Gupta +1
cs.CVcs.LGcs.ROarXiv:2007.00643v22020NeST: Neuron Selective Tuning for LLM Safety
Sasha Behrouzi, Lichao Wu, Mohamadreza Rostami +1
cs.CRcs.LGarXiv:2602.16835v22026BridgeData V2: A Dataset for Robot Learning at Scale
Homer Walke, Kevin Black, Abraham Lee +11
cs.ROcs.LGarXiv:2308.12952v32023Reinforced Attention Learning
Bangzheng Li, Jianmo Ni, Chen Qu +5
cs.CLcs.CVcs.LGarXiv:2602.04884v22026Self-Adaptive Physics-Informed Neural Networks using a Soft Attention Mechanism
Levi McClenny, Ulisses Braga-Neto
cs.LGstat.MLarXiv:2009.04544v52020MotionCrafter: Dense Geometry and Motion Reconstruction with a 4D VAE
Ruijie Zhu, Jiahao Lu, Wenbo Hu +4
cs.CVcs.AIcs.CGarXiv:2602.08961v22026Towards Actionable Surgical Team Dynamics: from Teamwork to Counterfactual Annotations
Vincenzo Marco De Luca, Antonio Longa, Andrea Passerini
cs.LGarXiv:2608.23344v12026Joint Optimization Framework for Learning with Noisy Labels
Daiki Tanaka, Daiki Ikami, Toshihiko Yamasaki +1
cs.CVcs.LGstat.MLarXiv:1803.11364v12018SMASH: One-Shot Model Architecture Search through HyperNetworks
Andrew Brock, Theodore Lim, J. M. Ritchie +1
cs.LGarXiv:1708.05344v12017Reasoning-Augmented Representations for Multimodal Retrieval
Jianrui Zhang, Anirudh Sundara Rajan, Brandon Han +3
cs.IRcs.AIcs.CVarXiv:2602.07125v12026Reinforcement Learning in Healthcare: A Survey
Chao Yu, Jiming Liu, Shamim Nemati
cs.LGcs.AIarXiv:1908.08796v42019Spectral Temporal Graph Neural Network for Multivariate Time-series Forecasting
Defu Cao, Yujing Wang, Juanyong Duan +8
cs.LGcs.AIarXiv:2103.07719v12021Stemphonic: All-at-once Flexible Multi-stem Music Generation
Shih-Lun Wu, Ge Zhu, Juan-Pablo Caceres +2
cs.SDcs.LGcs.MMarXiv:2602.09891v12026Hardware Co-Design Scaling Laws via Roofline Modelling for On-Device LLMs
Luoyang Sun, Jiwen Jiang, Yifeng Ding +9
cs.LGcs.CLarXiv:2602.10377v12026Structured Pruning of Deep Convolutional Neural Networks
Sajid Anwar, Kyuyeon Hwang, Wonyong Sung
cs.NEcs.LGstat.MLarXiv:1512.08571v12015A Survey of Inverse Reinforcement Learning: Challenges, Methods and Progress
Saurabh Arora, Prashant Doshi
cs.LGstat.MLarXiv:1806.06877v32018Video Frame Synthesis using Deep Voxel Flow
Ziwei Liu, Raymond A. Yeh, Xiaoou Tang +2
cs.CVcs.GRcs.LGarXiv:1702.02463v22017Multi-Task Learning with Deep Neural Networks: A Survey
Michael Crawshaw
cs.LGcs.CVstat.MLarXiv:2009.09796v12020UniT: Unified Multimodal Chain-of-Thought Test-time Scaling
Leon Liangyu Chen, Haoyu Ma, Zhipeng Fan +11
cs.CVcs.AIcs.LGarXiv:2602.12279v22026The (Un)reliability of saliency methods
Pieter-Jan Kindermans, Sara Hooker, Julius Adebayo +5
stat.MLcs.LGarXiv:1711.00867v12017DeepVision-103K: A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning
Haoxiang Sun, Lizhen Xu, Bing Zhao +5
cs.LGcs.AIarXiv:2602.16742v12026Human Pose Estimation with Iterative Error Feedback
Joao Carreira, Pulkit Agrawal, Katerina Fragkiadaki +1
cs.CVcs.LGcs.NEarXiv:1507.06550v32015Weight Decay Improves Language Model Plasticity
Tessa Han, Sebastian Bordt, Hanlin Zhang +1
cs.LGcs.AIcs.CLarXiv:2602.11137v22026Real-Time Flying Object Detection with YOLOv8
Dillon Reis, Jordan Kupec, Jacqueline Hong +1
cs.CVcs.LGarXiv:2305.09972v22023ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces
Xin Xu, Tong Yu, Xiang Chen +3
cs.AIcs.CLcs.LGarXiv:2602.11683v12026Variational Continual Learning
Cuong V. Nguyen, Yingzhen Li, Thang D. Bui +1
stat.MLcs.LGarXiv:1710.10628v32017Unsupervised Neural Machine Translation
Mikel Artetxe, Gorka Labaka, Eneko Agirre +1
cs.CLcs.AIcs.LGarXiv:1710.11041v22017Vector-quantized Image Modeling with Improved VQGAN
Jiahui Yu, Xin Li, Jing Yu Koh +7
cs.CVcs.LGarXiv:2110.04627v32021MAEB: Massive Audio Embedding Benchmark
Adnan El Assadi, Isaac Chung, Chenghao Xiao +15
cs.SDcs.AIcs.CLarXiv:2602.16008v12026Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Yunfei Chu, Jin Xu, Xiaohuan Zhou +5
eess.AScs.CLcs.LGarXiv:2311.07919v22023Using millions of emoji occurrences to learn any-domain representations for detecting sentiment, emotion and sarcasm
Bjarke Felbo, Alan Mislove, Anders Søgaard +2
stat.MLcs.LGarXiv:1708.00524v22017Linear Mode Connectivity and the Lottery Ticket Hypothesis
Jonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy +1
cs.LGcs.NEstat.MLarXiv:1912.05671v42019Sink-Aware Pruning for Diffusion Language Models
Aidar Myrzakhan, Tianyi Li, Bowei Guo +2
cs.CLcs.AIcs.LGarXiv:2602.17664v12026Reinforcement Learning for Combinatorial Optimization: A Survey
Nina Mazyavkina, Sergey Sviridov, Sergei Ivanov +1
cs.LGmath.COmath.OCarXiv:2003.03600v32020Synthetic and Natural Noise Both Break Neural Machine Translation
Yonatan Belinkov, Yonatan Bisk
cs.CLcs.LGarXiv:1711.02173v22017Beyond the Harness: End-to-End Optimization of Context Artifacts for Enterprise Text-to-SQL
Kate Gwimm, Carson Eisenach
cs.AIcs.LGarXiv:2608.22830v12026code2seq: Generating Sequences from Structured Representations of Code
Uri Alon, Shaked Brody, Omer Levy +1
cs.LGcs.PLstat.MLarXiv:1808.01400v62018ExpeL: LLM Agents Are Experiential Learners
Andrew Zhao, Daniel Huang, Quentin Xu +3
cs.LGcs.AIcs.CLarXiv:2308.10144v32023On the "Induction Bias" in Sequence Models
M. Reza Ebrahimi, Michaël Defferrard, Sunny Panchal +1
cs.LGcs.CLarXiv:2602.18333v22026Nesterov Accelerated Gradient and Scale Invariance for Adversarial Attacks
Jiadong Lin, Chuanbiao Song, Kun He +2
cs.LGcs.CRstat.MLarXiv:1908.06281v52019LFPO: Likelihood-Free Policy Optimization for Masked Diffusion Models
Chenxing Wei, Jiazhen Kang, Hong Wang +8
cs.LGcs.AIarXiv:2603.01563v12026Transition-Based Dependency Parsing with Stack Long Short-Term Memory
Chris Dyer, Miguel Ballesteros, Wang Ling +2
cs.CLcs.LGcs.NEarXiv:1505.08075v12015MMD GAN: Towards Deeper Understanding of Moment Matching Network
Chun-Liang Li, Wei-Cheng Chang, Yu Cheng +2
cs.LGcs.AIstat.MLarXiv:1705.08584v32017Test-Time Adaptation for ECG Classification via SQI-Gated Self-Training and Beat-Rhythm Consistency
Wenhan Jiang, Zhipeng Deng, Jiale Zhou +3
cs.LGarXiv:2608.23347v12026Virtual-to-real Deep Reinforcement Learning: Continuous Control of Mobile Robots for Mapless Navigation
Lei Tai, Giuseppe Paolo, Ming Liu
cs.ROcs.AIcs.LGarXiv:1703.00420v42017SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching
Yasaman Haghighi, Alexandre Alahi
cs.CVcs.LGarXiv:2602.24208v12026PTE: Predictive Text Embedding through Large-scale Heterogeneous Text Networks
Jian Tang, Meng Qu, Qiaozhu Mei
cs.CLcs.LGcs.NEarXiv:1508.00200v12015A Brief Review of Domain Adaptation
Abolfazl Farahani, Sahar Voghoei, Khaled Rasheed +1
cs.LGcs.CVarXiv:2010.03978v12020Agentic Critical Training
Weize Liu, Minghui Liu, Sy-Tuyen Ho +3
cs.AIcs.CLcs.LGarXiv:2603.08706v12026Consistency of the group Lasso and multiple kernel learning
Francis Bach
cs.LGarXiv:0707.3390v22007MetaCaster: Meta-Harness-Optimized Agent for End-to-End Few-Shot Learning of Lightweight Time Series Forecasters
ChengAo Shen, Wenchao Yu, Fangyu Wu +6
cs.LGcs.AIarXiv:2608.23473v12026Human-level performance in first-person multiplayer games with population-based deep reinforcement learning
Max Jaderberg, Wojciech M. Czarnecki, Iain Dunning +15
cs.LGcs.AIstat.MLarXiv:1807.01281v12018How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities
Ziwen Xu, Kewei Xu, Haoming Xu +8
cs.CLcs.AIcs.HCarXiv:2603.02578v22026Random Search and Reproducibility for Neural Architecture Search
Liam Li, Ameet Talwalkar
cs.LGstat.MLarXiv:1902.07638v32019TourPlanner: A Competitive Consensus Framework with Constraint-Gated Reinforcement Learning for Travel Planning
Yinuo Wang, Mining Tan, Wenxiang Jiao +5
cs.AIcs.CLcs.LGarXiv:2601.04698v12026