Machine Learning
Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,681 to 1,740 of 20,193
LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models
Mingyang Xie, Numair Khan, Tianfu Wang +8
cs.CVcs.LGarXiv:2601.14674v22026HittER: Hierarchical Transformers for Knowledge Graph Embeddings
Sanxing Chen, Xiaodong Liu, Jianfeng Gao +3
cs.CLcs.LGarXiv:2008.12813v22020Likely to stop? Predicting Stopout in Massive Open Online Courses
Colin Taylor, Kalyan Veeramachaneni, Una-May O'Reilly
cs.CYcs.LGarXiv:1408.3382v12014Does Neural Machine Translation Benefit from Larger Context?
Sebastien Jean, Stanislas Lauly, Orhan Firat +1
stat.MLcs.CLcs.LGarXiv:1704.05135v12017TableFormer: Table Structure Understanding with Transformers
Ahmed Nassar, Nikolaos Livathinos, Maksym Lysak +1
cs.CVcs.LGarXiv:2203.01017v22022ImageCAS: A Large-Scale Dataset and Benchmark for Coronary Artery Segmentation based on Computed Tomography Angiography Images
An Zeng, Chunbiao Wu, Meiping Huang +10
eess.IVcs.LGarXiv:2211.01607v22022HyperImpute: Generalized Iterative Imputation with Automatic Model Selection
Daniel Jarrett, Bogdan Cebere, Tennison Liu +2
stat.MLcs.LGarXiv:2206.07769v12022Combinatorial Multi-Armed Bandit with General Reward Functions
Wei Chen, Wei Hu, Fu Li +3
cs.LGcs.DSstat.MLarXiv:1610.06603v42016Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms
Michael Hanna, Sandro Pezzelle, Yonatan Belinkov
cs.LGcs.CLarXiv:2403.17806v22024How Useful is Self-Supervised Pretraining for Visual Tasks?
Alejandro Newell, Jia Deng
cs.CVcs.LGarXiv:2003.14323v12020Structured Graph Learning for Clustering and Semi-supervised Classification
Zhao Kang, Chong Peng, Qiang Cheng +4
cs.LGcs.AIcs.CVarXiv:2008.13429v12020Extreme Gradient Boosting for Yield Estimation compared with Deep Learning Approaches
Florian Huber, Artem Yushchenko, Benedikt Stratmann +1
cs.LGarXiv:2208.12633v12022TFAD: A Decomposition Time Series Anomaly Detection Architecture with Time-Frequency Analysis
Chaoli Zhang, Tian Zhou, Qingsong Wen +1
cs.LGcs.AIarXiv:2210.09693v22022Self-supervised Knowledge Distillation Using Singular Value Decomposition
Seung Hyun Lee, Dae Ha Kim, Byung Cheol Song
cs.LGcs.CVstat.MLarXiv:1807.06819v12018Rebalancing Token Importance in Language Models with TF-IDF Weighted Cross-Entropy Loss
Zhijian Li, Stefan Larson, Kevin Leach
cs.CLcs.LGarXiv:2609.11029v12026Time Series Change Point Detection with Self-Supervised Contrastive Predictive Coding
Shohreh Deldari, Daniel V. Smith, Hao Xue +1
cs.LGcs.AIcs.CVarXiv:2011.14097v52020Model-based Exploration of the Frontier of Behaviours for Deep Learning System Testing
Vincenzo Riccio, Paolo Tonella
cs.SEcs.AIcs.LGarXiv:2007.02787v12020DMD: A Large-Scale Multi-Modal Driver Monitoring Dataset for Attention and Alertness Analysis
Juan Diego Ortega, Neslihan Kose, Paola Cañas +5
cs.CVcs.LGeess.IVarXiv:2008.12085v12020Influence-Preserving Proxies for Gradient-Based Data Selection in LLM Fine-tuning
Sirui Chen, Yunzhe Qi, Mengting Ai +4
cs.LGarXiv:2602.17835v12026PEER: A Comprehensive and Multi-Task Benchmark for Protein Sequence Understanding
Minghao Xu, Zuobai Zhang, Jiarui Lu +5
cs.LGarXiv:2206.02096v22022A Positive Case for Faithfulness: LLM Self-Explanations Help Predict Model Behavior
Harry Mayne, Justin Singh Kang, Dewi Gould +3
cs.AIcs.LGarXiv:2602.02639v12026When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text
DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang +1
cs.CLcs.LGarXiv:2609.11067v12026Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions
Reem I. Masoud, Ziquan Liu, Martin Ferianc +2
cs.CYcs.CLcs.LGarXiv:2309.12342v22023Dynamic Graph Representation Learning via Self-Attention Networks
Aravind Sankar, Yanhong Wu, Liang Gou +2
cs.LGcs.SIstat.MLarXiv:1812.09430v22018Graph Anomaly Detection with Graph Neural Networks: Current Status and Challenges
Hwan Kim, Byung Suk Lee, Won-Yong Shin +1
cs.LGcs.AIcs.SIarXiv:2209.14930v22022Adversarial Distributional Training for Robust Deep Learning
Yinpeng Dong, Zhijie Deng, Tianyu Pang +2
cs.LGcs.CRstat.MLarXiv:2002.05999v22020PIGNet: A physics-informed deep learning model toward generalized drug-target interaction predictions
Seokhyun Moon, Wonho Zhung, Soojung Yang +2
q-bio.BMcs.LGarXiv:2008.12249v22020Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations
Kang Min Yoo, Junyeob Kim, Hyuhng Joon Kim +5
cs.CLcs.AIcs.LGarXiv:2205.12685v22022LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models
Ahmad Faiz, Sotaro Kaneda, Ruhan Wang +4
cs.CLcs.AIcs.CYarXiv:2309.14393v22023Robust Multimodal Sentiment Analysis with Incomplete Modalities via Semantic-aware Completeness based Reconstruction
Han-Jun Choi, Byunggill Joe, Saim Shin +1
cs.CLcs.AIcs.LGarXiv:2609.10950v12026Theoretical Guarantees for Permutation-Equivariant Quantum Neural Networks
Louis Schatzki, Martin Larocca, Quynh T. Nguyen +2
quant-phcs.LGstat.MLarXiv:2210.09974v32022HeurekaBench: A Benchmarking Framework for AI Co-scientist
Siba Smarak Panigrahi, Jovana Videnović, Maria Brbić
cs.LGarXiv:2601.01678v22026Structurally Speaking: Motif-Oriented Graph Captioning through Bidirectional Graph-Text Translation
Hsiao-Ying Lu, Dongyu Liu, Kwan-Liu Ma
cs.CLcs.LGarXiv:2609.10923v12026Flexible Job Shop Scheduling via Dual Attention Network Based Reinforcement Learning
Runqing Wang, Gang Wang, Jian Sun +2
cs.LGcs.AIarXiv:2305.05119v22023Virchow: A Million-Slide Digital Pathology Foundation Model
Eugene Vorontsov, Alican Bozkurt, Adam Casson +28
eess.IVcs.CVcs.LGarXiv:2309.07778v62023Discovering Causal Relations and Equations from Data
Gustau Camps-Valls, Andreas Gerhardus, Urmi Ninad +7
physics.data-ancs.AIcs.LGarXiv:2305.13341v12023Cross-Architecture Model Diffing with Crosscoders: Unsupervised Discovery of Differences Between LLMs
Thomas Jiralerspong, Trenton Bricken
cs.AIcs.LGcs.SEarXiv:2602.11729v12026CDRRM: Contrast-Driven Rubric Generation for Reliable and Interpretable Reward Modeling
Dengcan Liu, Fengkai Yang, Xiaohan Wang +7
cs.AIcs.LGarXiv:2603.08035v12026NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks
Swaroop Mishra, Arindam Mitra, Neeraj Varshney +4
cs.CLcs.AIcs.LGarXiv:2204.05660v12022Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark
Yihua Zhang, Pingzhi Li, Junyuan Hong +10
cs.LGcs.CLarXiv:2402.11592v32024Invertible Concept-based Explanations for CNN Models with Non-negative Concept Activation Vectors
Ruihan Zhang, Prashan Madumal, Tim Miller +2
cs.CVcs.AIcs.LGarXiv:2006.15417v42020Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
Jiashu Xu, Mingyu Derek Ma, Fei Wang +2
cs.CLcs.AIcs.CRarXiv:2305.14710v22023Characterizing and overcoming the greedy nature of learning in multi-modal deep neural networks
Nan Wu, Stanisław Jastrzębski, Kyunghyun Cho +1
cs.LGcs.CVarXiv:2202.05306v32022CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think
Junzhe Shen, Jieru Zhao, Ziwei He +1
cs.CLcs.AIcs.LGarXiv:2603.02547v12026Online stochastic gradient descent on non-convex losses from high-dimensional inference
Gerard Ben Arous, Reza Gheissari, Aukosh Jagannath
stat.MLcs.LGmath.PRarXiv:2003.10409v42020Detectable Only Where It Is Confounded: What Verified Duplication Counts Say About Membership Evidence in Language Models
Arman Nik Khah
cs.CLcs.CRcs.LGarXiv:2609.10830v12026Automatic Crack Detection on Road Pavements Using Encoder Decoder Architecture
Zhun Fan, Chong Li, Ying Chen +4
cs.CVcs.LGeess.IVarXiv:2007.00477v12020Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu
Farah Adeeba, Abdul Rafae Khan, Rajesh Bhatt +1
cs.CLcs.AIcs.LGarXiv:2609.10758v12026GCGNet: Graph-Consistent Generative Network for Time Series Forecasting with Exogenous Variables
Zhengyu Li, Xiangfei Qiu, Yuhan Zhu +4
cs.LGcs.AIarXiv:2603.08032v22026Multivariate Confidence Calibration for Object Detection
Fabian Küppers, Jan Kronenberger, Amirhossein Shantia +1
cs.CVcs.LGstat.MLarXiv:2004.13546v12020Radiomics in Medical Imaging: Methods, Applications, and Challenges
Fnu Neha, Deepak kumar Shukla
eess.IVcs.AIcs.LGarXiv:2602.00102v12026Neural Networks for Entity Matching: A Survey
Nils Barlaug, Jon Atle Gulla
cs.DBcs.CLcs.LGarXiv:2010.11075v22020A Comprehensive Survey on Data Augmentation
Zaitian Wang, Pengfei Wang, Kunpeng Liu +6
cs.LGcs.AIarXiv:2405.09591v42024Fourier-DeepONet: Fourier-enhanced deep operator networks for full waveform inversion with improved accuracy, generalizability, and robustness
Min Zhu, Shihang Feng, Youzuo Lin +1
cs.LGphysics.comp-phphysics.geo-pharXiv:2305.17289v22023A review on data-driven constitutive laws for solids
Jan Niklas Fuhg, Govinda Anantha Padmanabha, Nikolaos Bouklas +6
cs.CEcs.LGphysics.app-pharXiv:2405.03658v12024Polymer Informatics with Multi-Task Learning
Christopher Künneth, Arunkumar Chitteth Rajan, Huan Tran +3
cond-mat.mtrl-scics.LGphysics.comp-pharXiv:2010.15166v12020FedDisco: Federated Learning with Discrepancy-Aware Collaboration
Rui Ye, Mingkai Xu, Jianyu Wang +3
cs.LGarXiv:2305.19229v12023Diffusion Models Beat GANs on Topology Optimization
François Mazé, Faez Ahmed
cs.LGcs.CEarXiv:2208.09591v22022HIQL: Offline Goal-Conditioned RL with Latent States as Actions
Seohong Park, Dibya Ghosh, Benjamin Eysenbach +1
cs.LGcs.AIcs.ROarXiv:2307.11949v42023Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics
Chenhao Li, Andreas Krause, Marco Hutter
cs.ROcs.AIcs.LGarXiv:2501.10100v52025