Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,681 to 1,740 of 20,193

  1. LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models

    Mingyang Xie, Numair Khan, Tianfu Wang +8

    cs.CVcs.LGarXiv:2601.14674v22026
  2. HittER: Hierarchical Transformers for Knowledge Graph Embeddings

    Sanxing Chen, Xiaodong Liu, Jianfeng Gao +3

    cs.CLcs.LGarXiv:2008.12813v22020
  3. Likely to stop? Predicting Stopout in Massive Open Online Courses

    Colin Taylor, Kalyan Veeramachaneni, Una-May O'Reilly

    cs.CYcs.LGarXiv:1408.3382v12014
  4. Does Neural Machine Translation Benefit from Larger Context?

    Sebastien Jean, Stanislas Lauly, Orhan Firat +1

    stat.MLcs.CLcs.LGarXiv:1704.05135v12017
  5. TableFormer: Table Structure Understanding with Transformers

    Ahmed Nassar, Nikolaos Livathinos, Maksym Lysak +1

    cs.CVcs.LGarXiv:2203.01017v22022
  6. ImageCAS: A Large-Scale Dataset and Benchmark for Coronary Artery Segmentation based on Computed Tomography Angiography Images

    An Zeng, Chunbiao Wu, Meiping Huang +10

    eess.IVcs.LGarXiv:2211.01607v22022
  7. HyperImpute: Generalized Iterative Imputation with Automatic Model Selection

    Daniel Jarrett, Bogdan Cebere, Tennison Liu +2

    stat.MLcs.LGarXiv:2206.07769v12022
  8. Combinatorial Multi-Armed Bandit with General Reward Functions

    Wei Chen, Wei Hu, Fu Li +3

    cs.LGcs.DSstat.MLarXiv:1610.06603v42016
  9. Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms

    Michael Hanna, Sandro Pezzelle, Yonatan Belinkov

    cs.LGcs.CLarXiv:2403.17806v22024
  10. How Useful is Self-Supervised Pretraining for Visual Tasks?

    Alejandro Newell, Jia Deng

    cs.CVcs.LGarXiv:2003.14323v12020
  11. Structured Graph Learning for Clustering and Semi-supervised Classification

    Zhao Kang, Chong Peng, Qiang Cheng +4

    cs.LGcs.AIcs.CVarXiv:2008.13429v12020
  12. Extreme Gradient Boosting for Yield Estimation compared with Deep Learning Approaches

    Florian Huber, Artem Yushchenko, Benedikt Stratmann +1

    cs.LGarXiv:2208.12633v12022
  13. TFAD: A Decomposition Time Series Anomaly Detection Architecture with Time-Frequency Analysis

    Chaoli Zhang, Tian Zhou, Qingsong Wen +1

    cs.LGcs.AIarXiv:2210.09693v22022
  14. Self-supervised Knowledge Distillation Using Singular Value Decomposition

    Seung Hyun Lee, Dae Ha Kim, Byung Cheol Song

    cs.LGcs.CVstat.MLarXiv:1807.06819v12018
  15. Rebalancing Token Importance in Language Models with TF-IDF Weighted Cross-Entropy Loss

    Zhijian Li, Stefan Larson, Kevin Leach

    cs.CLcs.LGarXiv:2609.11029v12026
  16. Time Series Change Point Detection with Self-Supervised Contrastive Predictive Coding

    Shohreh Deldari, Daniel V. Smith, Hao Xue +1

    cs.LGcs.AIcs.CVarXiv:2011.14097v52020
  17. Model-based Exploration of the Frontier of Behaviours for Deep Learning System Testing

    Vincenzo Riccio, Paolo Tonella

    cs.SEcs.AIcs.LGarXiv:2007.02787v12020
  18. DMD: A Large-Scale Multi-Modal Driver Monitoring Dataset for Attention and Alertness Analysis

    Juan Diego Ortega, Neslihan Kose, Paola Cañas +5

    cs.CVcs.LGeess.IVarXiv:2008.12085v12020
  19. Influence-Preserving Proxies for Gradient-Based Data Selection in LLM Fine-tuning

    Sirui Chen, Yunzhe Qi, Mengting Ai +4

    cs.LGarXiv:2602.17835v12026
  20. PEER: A Comprehensive and Multi-Task Benchmark for Protein Sequence Understanding

    Minghao Xu, Zuobai Zhang, Jiarui Lu +5

    cs.LGarXiv:2206.02096v22022
  21. A Positive Case for Faithfulness: LLM Self-Explanations Help Predict Model Behavior

    Harry Mayne, Justin Singh Kang, Dewi Gould +3

    cs.AIcs.LGarXiv:2602.02639v12026
  22. When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text

    DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang +1

    cs.CLcs.LGarXiv:2609.11067v12026
  23. Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions

    Reem I. Masoud, Ziquan Liu, Martin Ferianc +2

    cs.CYcs.CLcs.LGarXiv:2309.12342v22023
  24. Dynamic Graph Representation Learning via Self-Attention Networks

    Aravind Sankar, Yanhong Wu, Liang Gou +2

    cs.LGcs.SIstat.MLarXiv:1812.09430v22018
  25. Graph Anomaly Detection with Graph Neural Networks: Current Status and Challenges

    Hwan Kim, Byung Suk Lee, Won-Yong Shin +1

    cs.LGcs.AIcs.SIarXiv:2209.14930v22022
  26. Adversarial Distributional Training for Robust Deep Learning

    Yinpeng Dong, Zhijie Deng, Tianyu Pang +2

    cs.LGcs.CRstat.MLarXiv:2002.05999v22020
  27. PIGNet: A physics-informed deep learning model toward generalized drug-target interaction predictions

    Seokhyun Moon, Wonho Zhung, Soojung Yang +2

    q-bio.BMcs.LGarXiv:2008.12249v22020
  28. Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations

    Kang Min Yoo, Junyeob Kim, Hyuhng Joon Kim +5

    cs.CLcs.AIcs.LGarXiv:2205.12685v22022
  29. LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models

    Ahmad Faiz, Sotaro Kaneda, Ruhan Wang +4

    cs.CLcs.AIcs.CYarXiv:2309.14393v22023
  30. Robust Multimodal Sentiment Analysis with Incomplete Modalities via Semantic-aware Completeness based Reconstruction

    Han-Jun Choi, Byunggill Joe, Saim Shin +1

    cs.CLcs.AIcs.LGarXiv:2609.10950v12026
  31. Theoretical Guarantees for Permutation-Equivariant Quantum Neural Networks

    Louis Schatzki, Martin Larocca, Quynh T. Nguyen +2

    quant-phcs.LGstat.MLarXiv:2210.09974v32022
  32. HeurekaBench: A Benchmarking Framework for AI Co-scientist

    Siba Smarak Panigrahi, Jovana Videnović, Maria Brbić

    cs.LGarXiv:2601.01678v22026
  33. Structurally Speaking: Motif-Oriented Graph Captioning through Bidirectional Graph-Text Translation

    Hsiao-Ying Lu, Dongyu Liu, Kwan-Liu Ma

    cs.CLcs.LGarXiv:2609.10923v12026
  34. Flexible Job Shop Scheduling via Dual Attention Network Based Reinforcement Learning

    Runqing Wang, Gang Wang, Jian Sun +2

    cs.LGcs.AIarXiv:2305.05119v22023
  35. Virchow: A Million-Slide Digital Pathology Foundation Model

    Eugene Vorontsov, Alican Bozkurt, Adam Casson +28

    eess.IVcs.CVcs.LGarXiv:2309.07778v62023
  36. Discovering Causal Relations and Equations from Data

    Gustau Camps-Valls, Andreas Gerhardus, Urmi Ninad +7

    physics.data-ancs.AIcs.LGarXiv:2305.13341v12023
  37. Cross-Architecture Model Diffing with Crosscoders: Unsupervised Discovery of Differences Between LLMs

    Thomas Jiralerspong, Trenton Bricken

    cs.AIcs.LGcs.SEarXiv:2602.11729v12026
  38. CDRRM: Contrast-Driven Rubric Generation for Reliable and Interpretable Reward Modeling

    Dengcan Liu, Fengkai Yang, Xiaohan Wang +7

    cs.AIcs.LGarXiv:2603.08035v12026
  39. NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks

    Swaroop Mishra, Arindam Mitra, Neeraj Varshney +4

    cs.CLcs.AIcs.LGarXiv:2204.05660v12022
  40. Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark

    Yihua Zhang, Pingzhi Li, Junyuan Hong +10

    cs.LGcs.CLarXiv:2402.11592v32024
  41. Invertible Concept-based Explanations for CNN Models with Non-negative Concept Activation Vectors

    Ruihan Zhang, Prashan Madumal, Tim Miller +2

    cs.CVcs.AIcs.LGarXiv:2006.15417v42020
  42. Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models

    Jiashu Xu, Mingyu Derek Ma, Fei Wang +2

    cs.CLcs.AIcs.CRarXiv:2305.14710v22023
  43. Characterizing and overcoming the greedy nature of learning in multi-modal deep neural networks

    Nan Wu, Stanisław Jastrzębski, Kyunghyun Cho +1

    cs.LGcs.CVarXiv:2202.05306v32022
  44. CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think

    Junzhe Shen, Jieru Zhao, Ziwei He +1

    cs.CLcs.AIcs.LGarXiv:2603.02547v12026
  45. Online stochastic gradient descent on non-convex losses from high-dimensional inference

    Gerard Ben Arous, Reza Gheissari, Aukosh Jagannath

    stat.MLcs.LGmath.PRarXiv:2003.10409v42020
  46. Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models

    Arman Nik Khah

    cs.CLcs.CRcs.LGarXiv:2609.10830v12026
  47. Automatic Crack Detection on Road Pavements Using Encoder Decoder Architecture

    Zhun Fan, Chong Li, Ying Chen +4

    cs.CVcs.LGeess.IVarXiv:2007.00477v12020
  48. Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu

    Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt +1

    cs.CLcs.AIcs.LGarXiv:2609.10758v12026
  49. GCGNet: Graph-Consistent Generative Network for Time Series Forecasting with Exogenous Variables

    Zhengyu Li, Xiangfei Qiu, Yuhan Zhu +4

    cs.LGcs.AIarXiv:2603.08032v22026
  50. Multivariate Confidence Calibration for Object Detection

    Fabian Küppers, Jan Kronenberger, Amirhossein Shantia +1

    cs.CVcs.LGstat.MLarXiv:2004.13546v12020
  51. Radiomics in Medical Imaging: Methods, Applications, and Challenges

    Fnu Neha, Deepak kumar Shukla

    eess.IVcs.AIcs.LGarXiv:2602.00102v12026
  52. Neural Networks for Entity Matching: A Survey

    Nils Barlaug, Jon Atle Gulla

    cs.DBcs.CLcs.LGarXiv:2010.11075v22020
  53. A Comprehensive Survey on Data Augmentation

    Zaitian Wang, Pengfei Wang, Kunpeng Liu +6

    cs.LGcs.AIarXiv:2405.09591v42024
  54. Fourier-DeepONet: Fourier-enhanced deep operator networks for full waveform inversion with improved accuracy, generalizability, and robustness

    Min Zhu, Shihang Feng, Youzuo Lin +1

    cs.LGphysics.comp-phphysics.geo-pharXiv:2305.17289v22023
  55. A review on data-driven constitutive laws for solids

    Jan Niklas Fuhg, Govinda Anantha Padmanabha, Nikolaos Bouklas +6

    cs.CEcs.LGphysics.app-pharXiv:2405.03658v12024
  56. Polymer Informatics with Multi-Task Learning

    Christopher Künneth, Arunkumar Chitteth Rajan, Huan Tran +3

    cond-mat.mtrl-scics.LGphysics.comp-pharXiv:2010.15166v12020
  57. FedDisco: Federated Learning with Discrepancy-Aware Collaboration

    Rui Ye, Mingkai Xu, Jianyu Wang +3

    cs.LGarXiv:2305.19229v12023
  58. Diffusion Models Beat GANs on Topology Optimization

    François Mazé, Faez Ahmed

    cs.LGcs.CEarXiv:2208.09591v22022
  59. HIQL: Offline Goal-Conditioned RL with Latent States as Actions

    Seohong Park, Dibya Ghosh, Benjamin Eysenbach +1

    cs.LGcs.AIcs.ROarXiv:2307.11949v42023
  60. Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics

    Chenhao Li, Andreas Krause, Marco Hutter

    cs.ROcs.AIcs.LGarXiv:2501.10100v52025