Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
15,421 to 15,480 of 20,199
Fake News Detection on Social Media using Geometric Deep Learning
Federico Monti, Fabrizio Frasca, Davide Eynard +2
cs.SIcs.LGstat.MLarXiv:1902.06673v12019LESS: Selecting Influential Data for Targeted Instruction Tuning
Mengzhou Xia, Sadhika Malladi, Suchin Gururangan +2
cs.CLcs.AIcs.LGarXiv:2402.04333v32024FlowNet3D: Learning Scene Flow in 3D Point Clouds
Xingyu Liu, Charles R. Qi, Leonidas J. Guibas
cs.CVcs.LGarXiv:1806.01411v32018Top-K Off-Policy Correction for a REINFORCE Recommender System
Minmin Chen, Alex Beutel, Paul Covington +3
cs.LGcs.IRstat.MLarXiv:1812.02353v32018Tell Me Where to Look: Guided Attention Inference Network
Kunpeng Li, Ziyan Wu, Kuan-Chuan Peng +2
cs.CVcs.LGarXiv:1802.10171v12018Sim-to-Real Robot Learning from Pixels with Progressive Nets
Andrei A. Rusu, Mel Vecerik, Thomas Rothörl +3
cs.ROcs.LGarXiv:1610.04286v22016Found in Translation: Learning Robust Joint Representations by Cyclic Translations Between Modalities
Hai Pham, Paul Pu Liang, Thomas Manzini +2
cs.LGcs.CLcs.CVarXiv:1812.07809v22018NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models
Chankyu Lee, Rajarshi Roy, Mengyao Xu +4
cs.CLcs.AIcs.IRarXiv:2405.17428v32024Reinforcement Knowledge Graph Reasoning for Explainable Recommendation
Yikun Xian, Zuohui Fu, S. Muthukrishnan +2
cs.IRcs.LGarXiv:1906.05237v12019NiftyNet: a deep-learning platform for medical imaging
Eli Gibson, Wenqi Li, Carole Sudre +14
cs.CVcs.LGcs.NEarXiv:1709.03485v22017Detection != Reliable Control: Decodable Empathy Directions Yield at Most Partial Shifts in Automated Empathy Scores
Haoran Jisun
cs.CLcs.HCcs.LGarXiv:2608.24901v12026The geometry of AI validation: Exact certification limits for iid best-of-N search
Ricardo Fitas
cs.LGmath.STstat.MLarXiv:2608.21496v12026Insights on representational similarity in neural networks with canonical correlation
Ari S. Morcos, Maithra Raghu, Samy Bengio
stat.MLcs.AIcs.CVarXiv:1806.05759v32018StateTune: Transforming LLM-Assisted EDA Flow Tuning into a Stateful, Closed-Loop Process
Kunlong Li, Shangshang Yao, Su Zheng +1
cs.ARcs.LGcs.MAarXiv:2608.23601v12026Objects that Sound
Relja Arandjelović, Andrew Zisserman
cs.CVcs.LGcs.MMarXiv:1712.06651v22017The Plan, Not the Decoder: Diagnosing and Repairing Compositional Failure in Reasoning-Augmented Text-to-Image Generation
Ashritha Gonuguntla
cs.CVcs.CLcs.LGarXiv:2608.21713v12026Implicit Functions in Feature Space for 3D Shape Reconstruction and Completion
Julian Chibane, Thiemo Alldieck, Gerard Pons-Moll
cs.CVcs.LGarXiv:2003.01456v22020RiskWorld: Object-Centric Latent World Modeling for Autonomous Driving Risk Identification
Jingzheng Li, Yufei Ge, Qianren Mao +5
cs.ROcs.LGarXiv:2608.21414v12026Constructing Predictive Surgical Path for AI-based Capsulorhexis Skill Transfer
Mohammad Javad Ahmadi, Hamid D. Taghirad
cs.ROcs.AIcs.LGarXiv:2608.21441v12026Loss-Parameterized Fisher Width Along Learning Trajectories
Vu Khac Ky
cs.LGmath.STarXiv:2608.21561v12026Variational Structure at the Edge of Stability
Eric Regis
cs.LGarXiv:2608.21660v12026On the Optimization of Deep Networks: Implicit Acceleration by Overparameterization
Sanjeev Arora, Nadav Cohen, Elad Hazan
cs.LGarXiv:1802.06509v22018TANGO: Token-Aggregated Nonlinear Gating Operators for Natural and Formal Language Modeling
Joshua Nunley
cs.LGcs.CLarXiv:2608.22117v12026Autonomous Cyber Defense: Real-Time Attack Detection and Mitigation in Software-Defined Networks Using Machine Learning
Alexandre Amaral, Fernando Moro, Ana Malheiro
cs.CRcs.LGarXiv:2608.22075v32026EditStream: A Unified Autoregressive Framework for Interactive Video Generation and Editing
Yuqian Zhou, Zhenghong Zhou, Zongze Wu +5
cs.CVcs.GRcs.HCarXiv:2608.21424v12026Deep Decentralized Multi-task Multi-Agent Reinforcement Learning under Partial Observability
Shayegan Omidshafiei, Jason Pazis, Christopher Amato +2
cs.LGcs.AIcs.MAarXiv:1703.06182v42017Robust Estimators in High Dimensions without the Computational Intractability
Ilias Diakonikolas, Gautam Kamath, Daniel Kane +3
cs.DScs.ITcs.LGarXiv:1604.06443v22016Sorting from Counterexamples
Noga Alon, Shay Moran, Shlomo Moran
cs.LGcs.CCcs.CGarXiv:2608.21579v12026BeTaL-GBI: Admission-Aware Benchmark Tuning and Full-Stack Verification of Geometric Belief Interfaces
Alvin Spivey, Yu Huang
cs.SEcs.CRcs.LGarXiv:2608.21503v12026Essentially No Barriers in Neural Network Energy Landscape
Felix Draxler, Kambis Veschgini, Manfred Salmhofer +1
stat.MLcs.AIcs.LGarXiv:1803.00885v52018Towards Universal Paraphrastic Sentence Embeddings
John Wieting, Mohit Bansal, Kevin Gimpel +1
cs.CLcs.LGarXiv:1511.08198v32015Beyond temperature scaling: Obtaining well-calibrated multiclass probabilities with Dirichlet calibration
Meelis Kull, Miquel Perello-Nieto, Markus Kängsepp +3
cs.LGstat.MLarXiv:1910.12656v12019Exploring Long-period Architectures: Four New Planet Candidates from Kepler with Periods >342 days
Matthew T. Hansen, Jason A. Dittmann
astro-ph.EPastro-ph.IMcs.LGarXiv:2608.23425v12026TorchIO: A Python library for efficient loading, preprocessing, augmentation and patch-based sampling of medical images in deep learning
Fernando Pérez-García, Rachel Sparks, Sébastien Ourselin
eess.IVcs.AIcs.CVarXiv:2003.04696v52020Geometric Structures on Graphs: a Holonomy-Based Discretization of Curvature
Hao Li, Yuhan Peng, Junwen Dong
cs.LGmath.DGarXiv:2608.22453v12026Unveiling the Depth-Performance Dilemma in Split-Federated Fine-tuning of LLMs
Hariharan Ramesh, Someshwaran Murugaiyan, Jyotikrishna Dass
cs.LGcs.CLcs.DCarXiv:2608.22188v12026Scaling Distributed Machine Learning with In-Network Aggregation
Amedeo Sapio, Marco Canini, Chen-Yu Ho +7
cs.DCcs.LGcs.NIarXiv:1903.06701v22019Leveraging Remote Traffic Data for Local Air Pollutant Estimation: A Scenario-Based Machine Learning Study Across London Monitoring Sites
Valeria Legaria-Santiago, Amadeo Arguelles, Magdalena Saldana-Perez +2
cs.LGarXiv:2608.23219v12026Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Abhishek Gupta, Vikash Kumar, Corey Lynch +2
cs.LGcs.ROstat.MLarXiv:1910.11956v12019Complexity Induction: Compositional Generalization via Structured Label Distortion
Aleksandr Abramov
cs.CVcs.AIcs.LGarXiv:2608.21464v12026Posterior Information Dynamics of Diffusion Models for Linear Inverse Problems
Xiangming Meng
cs.LGcs.ITeess.SParXiv:2608.21709v12026Guidance for Prior Change via Density Ratio Estimation
Yichen Zang, Song Liu, Jiun-Yi Lin
stat.MLcs.LGstat.MEarXiv:2608.21729v12026Counterfactual Quotient Models: Learning What Actions Change, Not What the World Does
Junlin Chen, Ruijie Wang, Jianxin Li
cs.LGarXiv:2608.22092v12026Semantic Reasoning Denoising: Correcting Language Model Reasoning with Semantic Operators
Yujiao Yang
cs.CLcs.AIcs.LGarXiv:2608.22090v12026On the Utility of Learning about Humans for Human-AI Coordination
Micah Carroll, Rohin Shah, Mark K. Ho +4
cs.LGcs.AIcs.HCarXiv:1910.05789v22019Spectral partitioning for $k$-block averaging kernels of finite Markov chains
Michael C. H. Choi, Youjia Wang
stat.MLcs.ITcs.LGarXiv:2608.21466v12026Rethinking Communication Metrics: How Should We Measure Meaning?
Niloofar Tavakolian, Hakimeh Purmehdi, Jungyeon Baek
cs.LGarXiv:2608.21626v12026Class-Conditioned Gaussian Mixture Modeling for Imbalanced Time Series Quantification
Md Shahriar Kabir, Mayesha Maliha R. Mithila, Anne H. H. Ngu +2
cs.LGarXiv:2608.21473v12026Anchoring Bias: A Persistent Fairness Backdoor Attack against MLLMs under Continual Learning
Yuyang Luo, Kai Shu
cs.LGcs.AIarXiv:2608.21577v12026Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale
Matthew Le, Apoorv Vyas, Bowen Shi +8
eess.AScs.CLcs.LGarXiv:2306.15687v22023Forecasting: theory and practice
Fotios Petropoulos, Daniele Apiletti, Vassilios Assimakopoulos +77
stat.APcs.LGecon.EMarXiv:2012.03854v42020TherMapNet Attention-Guided Runtime Full-Chip Thermal Map Prediction from Performance Metrics
Qin Gu, Chaofang Ma, Mingyu Yang +4
cs.ARcs.LGarXiv:2608.21887v12026SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning
Hanrui Wang, Zhekai Zhang, Song Han
cs.ARcs.AIcs.CLarXiv:2012.09852v32020Channel-wise Autoregressive Entropy Models for Learned Image Compression
David Minnen, Saurabh Singh
eess.IVcs.CVcs.ITarXiv:2007.08739v12020An efficient framework for learning sentence representations
Lajanugen Logeswaran, Honglak Lee
cs.CLcs.AIcs.LGarXiv:1803.02893v12018Tensor Seeks Layout: Formalizing Layout Selection for ML Compilers
Clemens Eisenhofer, Yuwen Jia, Daniel Kroening +1
cs.PLcs.DScs.LGarXiv:2608.21555v12026Variance Driven Exploration: A Provable and Efficient Methodology for Pure Exploration in Highly Stochastic Environments
Khang Luong, Nam Nguyen, Hoang Ta +2
cs.LGcs.AIstat.MLarXiv:2608.21995v12026Multi-attention Recurrent Network for Human Communication Comprehension
Amir Zadeh, Paul Pu Liang, Soujanya Poria +3
cs.AIcs.CLcs.LGarXiv:1802.00923v12018Globally Normalized Transition-Based Neural Networks
Daniel Andor, Chris Alberti, David Weiss +5
cs.CLcs.LGcs.NEarXiv:1603.06042v22016Congruence Decomposition with Neural Block Solvers for Large-Scale PCI Assignment
Yeqing Qiu, Chengpiao Huang, Ye Xue +7
cs.LGeess.SParXiv:2608.21485v12026