Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

15,481 to 15,540 of 20,247

  1. Insights on representational similarity in neural networks with canonical correlation

    Ari S. Morcos, Maithra Raghu, Samy Bengio

    stat.MLcs.AIcs.CVarXiv:1806.05759v32018
  2. StateTune: Transforming LLM-Assisted EDA Flow Tuning into a Stateful, Closed-Loop Process

    Kunlong Li, Shangshang Yao, Su Zheng +1

    cs.ARcs.LGcs.MAarXiv:2608.23601v12026
  3. Objects that Sound

    Relja Arandjelović, Andrew Zisserman

    cs.CVcs.LGcs.MMarXiv:1712.06651v22017
  4. The Plan, Not the Decoder: Diagnosing and Repairing Compositional Failure in Reasoning-Augmented Text-to-Image Generation

    Ashritha Gonuguntla

    cs.CVcs.CLcs.LGarXiv:2608.21713v12026
  5. Implicit Functions in Feature Space for 3D Shape Reconstruction and Completion

    Julian Chibane, Thiemo Alldieck, Gerard Pons-Moll

    cs.CVcs.LGarXiv:2003.01456v22020
  6. RiskWorld: Object-Centric Latent World Modeling for Autonomous Driving Risk Identification

    Jingzheng Li, Yufei Ge, Qianren Mao +5

    cs.ROcs.LGarXiv:2608.21414v12026
  7. Constructing Predictive Surgical Path for AI-based Capsulorhexis Skill Transfer

    Mohammad Javad Ahmadi, Hamid D. Taghirad

    cs.ROcs.AIcs.LGarXiv:2608.21441v12026
  8. Loss-Parameterized Fisher Width Along Learning Trajectories

    Vu Khac Ky

    cs.LGmath.STarXiv:2608.21561v12026
  9. Variational Structure at the Edge of Stability

    Eric Regis

    cs.LGarXiv:2608.21660v12026
  10. On the Optimization of Deep Networks: Implicit Acceleration by Overparameterization

    Sanjeev Arora, Nadav Cohen, Elad Hazan

    cs.LGarXiv:1802.06509v22018
  11. TANGO: Token-Aggregated Nonlinear Gating Operators for Natural and Formal Language Modeling

    Joshua Nunley

    cs.LGcs.CLarXiv:2608.22117v12026
  12. Autonomous Cyber Defense: Real-Time Attack Detection and Mitigation in Software-Defined Networks Using Machine Learning

    Alexandre Amaral, Fernando Moro, Ana Malheiro

    cs.CRcs.LGarXiv:2608.22075v32026
  13. EditStream: A Unified Autoregressive Framework for Interactive Video Generation and Editing

    Yuqian Zhou, Zhenghong Zhou, Zongze Wu +5

    cs.CVcs.GRcs.HCarXiv:2608.21424v12026
  14. Deep Decentralized Multi-task Multi-Agent Reinforcement Learning under Partial Observability

    Shayegan Omidshafiei, Jason Pazis, Christopher Amato +2

    cs.LGcs.AIcs.MAarXiv:1703.06182v42017
  15. Robust Estimators in High Dimensions without the Computational Intractability

    Ilias Diakonikolas, Gautam Kamath, Daniel Kane +3

    cs.DScs.ITcs.LGarXiv:1604.06443v22016
  16. Sorting from Counterexamples

    Noga Alon, Shay Moran, Shlomo Moran

    cs.LGcs.CCcs.CGarXiv:2608.21579v12026
  17. BeTaL-GBI: Admission-Aware Benchmark Tuning and Full-Stack Verification of Geometric Belief Interfaces

    Alvin Spivey, Yu Huang

    cs.SEcs.CRcs.LGarXiv:2608.21503v12026
  18. Essentially No Barriers in Neural Network Energy Landscape

    Felix Draxler, Kambis Veschgini, Manfred Salmhofer +1

    stat.MLcs.AIcs.LGarXiv:1803.00885v52018
  19. Towards Universal Paraphrastic Sentence Embeddings

    John Wieting, Mohit Bansal, Kevin Gimpel +1

    cs.CLcs.LGarXiv:1511.08198v32015
  20. Beyond temperature scaling: Obtaining well-calibrated multiclass probabilities with Dirichlet calibration

    Meelis Kull, Miquel Perello-Nieto, Markus Kängsepp +3

    cs.LGstat.MLarXiv:1910.12656v12019
  21. Exploring Long-period Architectures: Four New Planet Candidates from Kepler with Periods >342 days

    Matthew T. Hansen, Jason A. Dittmann

    astro-ph.EPastro-ph.IMcs.LGarXiv:2608.23425v12026
  22. TorchIO: A Python library for efficient loading, preprocessing, augmentation and patch-based sampling of medical images in deep learning

    Fernando Pérez-García, Rachel Sparks, Sébastien Ourselin

    eess.IVcs.AIcs.CVarXiv:2003.04696v52020
  23. Geometric Structures on Graphs: a Holonomy-Based Discretization of Curvature

    Hao Li, Yuhan Peng, Junwen Dong

    cs.LGmath.DGarXiv:2608.22453v12026
  24. Unveiling the Depth-Performance Dilemma in Split-Federated Fine-tuning of LLMs

    Hariharan Ramesh, Someshwaran Murugaiyan, Jyotikrishna Dass

    cs.LGcs.CLcs.DCarXiv:2608.22188v12026
  25. Scaling Distributed Machine Learning with In-Network Aggregation

    Amedeo Sapio, Marco Canini, Chen-Yu Ho +7

    cs.DCcs.LGcs.NIarXiv:1903.06701v22019
  26. Leveraging Remote Traffic Data for Local Air Pollutant Estimation: A Scenario-Based Machine Learning Study Across London Monitoring Sites

    Valeria Legaria-Santiago, Amadeo Arguelles, Magdalena Saldana-Perez +2

    cs.LGarXiv:2608.23219v12026
  27. Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning

    Abhishek Gupta, Vikash Kumar, Corey Lynch +2

    cs.LGcs.ROstat.MLarXiv:1910.11956v12019
  28. Complexity Induction: Compositional Generalization via Structured Label Distortion

    Aleksandr Abramov

    cs.CVcs.AIcs.LGarXiv:2608.21464v12026
  29. Posterior Information Dynamics of Diffusion Models for Linear Inverse Problems

    Xiangming Meng

    cs.LGcs.ITeess.SParXiv:2608.21709v12026
  30. Guidance for Prior Change via Density Ratio Estimation

    Yichen Zang, Song Liu, Jiun-Yi Lin

    stat.MLcs.LGstat.MEarXiv:2608.21729v12026
  31. Counterfactual Quotient Models: Learning What Actions Change, Not What the World Does

    Junlin Chen, Ruijie Wang, Jianxin Li

    cs.LGarXiv:2608.22092v12026
  32. Semantic Reasoning Denoising: Correcting Language Model Reasoning with Semantic Operators

    Yujiao Yang

    cs.CLcs.AIcs.LGarXiv:2608.22090v12026
  33. On the Utility of Learning about Humans for Human-AI Coordination

    Micah Carroll, Rohin Shah, Mark K. Ho +4

    cs.LGcs.AIcs.HCarXiv:1910.05789v22019
  34. Spectral partitioning for $k$-block averaging kernels of finite Markov chains

    Michael C. H. Choi, Youjia Wang

    stat.MLcs.ITcs.LGarXiv:2608.21466v12026
  35. Rethinking Communication Metrics: How Should We Measure Meaning?

    Niloofar Tavakolian, Hakimeh Purmehdi, Jungyeon Baek

    cs.LGarXiv:2608.21626v12026
  36. Class-Conditioned Gaussian Mixture Modeling for Imbalanced Time Series Quantification

    Md Shahriar Kabir, Mayesha Maliha R. Mithila, Anne H. H. Ngu +2

    cs.LGarXiv:2608.21473v12026
  37. Anchoring Bias: A Persistent Fairness Backdoor Attack against MLLMs under Continual Learning

    Yuyang Luo, Kai Shu

    cs.LGcs.AIarXiv:2608.21577v12026
  38. Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale

    Matthew Le, Apoorv Vyas, Bowen Shi +8

    eess.AScs.CLcs.LGarXiv:2306.15687v22023
  39. Forecasting: theory and practice

    Fotios Petropoulos, Daniele Apiletti, Vassilios Assimakopoulos +77

    stat.APcs.LGecon.EMarXiv:2012.03854v42020
  40. TherMapNet Attention-Guided Runtime Full-Chip Thermal Map Prediction from Performance Metrics

    Qin Gu, Chaofang Ma, Mingyu Yang +4

    cs.ARcs.LGarXiv:2608.21887v12026
  41. SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head Pruning

    Hanrui Wang, Zhekai Zhang, Song Han

    cs.ARcs.AIcs.CLarXiv:2012.09852v32020
  42. Channel-wise Autoregressive Entropy Models for Learned Image Compression

    David Minnen, Saurabh Singh

    eess.IVcs.CVcs.ITarXiv:2007.08739v12020
  43. An efficient framework for learning sentence representations

    Lajanugen Logeswaran, Honglak Lee

    cs.CLcs.AIcs.LGarXiv:1803.02893v12018
  44. Tensor Seeks Layout: Formalizing Layout Selection for ML Compilers

    Clemens Eisenhofer, Yuwen Jia, Daniel Kroening +1

    cs.PLcs.DScs.LGarXiv:2608.21555v12026
  45. Variance Driven Exploration: A Provable and Efficient Methodology for Pure Exploration in Highly Stochastic Environments

    Khang Luong, Nam Nguyen, Hoang Ta +2

    cs.LGcs.AIstat.MLarXiv:2608.21995v12026
  46. Multi-attention Recurrent Network for Human Communication Comprehension

    Amir Zadeh, Paul Pu Liang, Soujanya Poria +3

    cs.AIcs.CLcs.LGarXiv:1802.00923v12018
  47. Globally Normalized Transition-Based Neural Networks

    Daniel Andor, Chris Alberti, David Weiss +5

    cs.CLcs.LGcs.NEarXiv:1603.06042v22016
  48. Congruence Decomposition with Neural Block Solvers for Large-Scale PCI Assignment

    Yeqing Qiu, Chengpiao Huang, Ye Xue +7

    cs.LGeess.SParXiv:2608.21485v12026
  49. ChemDIRT: A Diversified Instruction, Representation, and Task Benchmark for Robust Chemistry-LLM Evaluation

    Eric Inae, Tim Gunn, Chris Bond +1

    cs.LGarXiv:2608.21504v12026
  50. Adversarial Variational Bayes: Unifying Variational Autoencoders and Generative Adversarial Networks

    Lars Mescheder, Sebastian Nowozin, Andreas Geiger

    cs.LGarXiv:1701.04722v42017
  51. Blended Latent Diffusion

    Omri Avrahami, Ohad Fried, Dani Lischinski

    cs.CVcs.GRcs.LGarXiv:2206.02779v22022
  52. Predicting Early Functional Decline from Longitudinal Laboratory and Vital Sign Trajectories: A Large-Scale Study Using the All of Us Research Program

    Rashmita Kudamala, Aravind V. Kuruvikkattil, Lalitha Pranathi Pulavarthy +1

    cs.LGstat.AParXiv:2608.21589v12026
  53. Guaranteed Rank Minimization via Singular Value Projection

    Raghu Meka, Prateek Jain, Inderjit S. Dhillon

    cs.LGcs.ITarXiv:0909.5457v32009
  54. Gauss--Hermite Quadrature for Gaussian-Mixture Entropy with an Action-Space Hermite Surrogate

    Jae Wan Shim

    stat.MLcs.ITcs.LGarXiv:2608.21467v22026
  55. GeoQ: Geometry-Aware Conditional Quantile Error Estimation for Scientific Surrogate Models

    Khoa Nguyen, Daniel Serino, Aviral Prakash +1

    cs.LGstat.MLarXiv:2608.21652v12026
  56. SynEHR: Joint Modeling Inter-visit Temporal Evolution and Intra-visit Clinical Structure for Longitudinal EHR Synthesis

    Ximiao Li, Lin Jiang, Rongchao Xu +3

    cs.LGcs.AIarXiv:2608.21673v12026
  57. Class Imbalance Problem in Data Mining Review

    Rushi Longadge, Snehalata Dongre

    cs.LGarXiv:1305.1707v12013
  58. Transfer Learning with Deep Convolutional Neural Network (CNN) for Pneumonia Detection using Chest X-ray

    Tawsifur Rahman, Muhammad E. H. Chowdhury, Amith Khandakar +5

    eess.IVcs.CVcs.LGarXiv:2004.06578v12020
  59. Learning to Prune Deep Neural Networks via Layer-wise Optimal Brain Surgeon

    Xin Dong, Shangyu Chen, Sinno Jialin Pan

    cs.NEcs.CVcs.LGarXiv:1705.07565v22017
  60. When More References Hurt: Contamination-Aware DINOv2 Memory Banks for Few-Shot Steel Defect Detection

    Hannaneh Kalantary, Javad Khoramdel

    cs.CVcs.LGarXiv:2608.22082v22026