Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,521 to 8,580 of 20,205

  1. Retrieved but not ranked: surface-form bias in structural retrieval, from mathematics to agent trajectories

    Nabira Rashid, Manolis Kellis

    cs.LGcs.AIcs.IRarXiv:2609.01556v12026
  2. World Model on Million-Length Video And Language With Blockwise RingAttention

    Hao Liu, Wilson Yan, Matei Zaharia +1

    cs.LGarXiv:2402.08268v42024
  3. Gradient Boosted Feature Selection

    Zhixiang Eddie Xu, Gao Huang, Kilian Q. Weinberger +1

    cs.LGstat.MLarXiv:1901.04055v12019
  4. PCONV: The Missing but Desirable Sparsity in DNN Weight Pruning for Real-time Execution on Mobile Devices

    Xiaolong Ma, Fu-Ming Guo, Wei Niu +5

    cs.LGcs.CVcs.DCarXiv:1909.05073v42019
  5. Prototype-guided transfer of sparse literature knowledge for electrolyte additive discovery

    Weixiang Hong, Hongting Du, Jiayue Tang +4

    physics.chem-phcs.LGarXiv:2609.02209v12026
  6. Masked Diffusion Models are Secretly Time-Agnostic Masked Models and Exploit Inaccurate Categorical Sampling

    Kaiwen Zheng, Yongxin Chen, Hanzi Mao +3

    cs.LGcs.AIcs.CLarXiv:2409.02908v62024
  7. AutoCompress: An Automatic DNN Structured Pruning Framework for Ultra-High Compression Rates

    Ning Liu, Xiaolong Ma, Zhiyuan Xu +3

    cs.LGcs.AIcs.CVarXiv:1907.03141v22019
  8. Learning to Generate Images with Perceptual Similarity Metrics

    Jake Snell, Karl Ridgeway, Renjie Liao +3

    cs.LGcs.CVarXiv:1511.06409v32015
  9. CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code Generation

    Varun Gadey, Ziad Marey, Alexandra Dmitrienko

    cs.CRcs.LGarXiv:2609.02774v12026
  10. TreeGen: A Tree-Based Transformer Architecture for Code Generation

    Zeyu Sun, Qihao Zhu, Yingfei Xiong +3

    cs.LGcs.SEarXiv:1911.09983v22019
  11. Principles and Algorithms for Forecasting Groups of Time Series: Locality and Globality

    Pablo Montero-Manso, Rob J Hyndman

    cs.LGstat.MLarXiv:2008.00444v32020
  12. GenCAR: Generative Counterfactual Alignment with Risk-Controlled Selection for Out-of-Distribution Recommendation

    Qianqian Wang, Yunshan Li, Jiawen Zeng +2

    cs.IRcs.LGarXiv:2609.02162v12026
  13. UE5M3 FP4 Block Scaling for Stable Language Model Pretraining

    Robert Hu, Carlo Luschi, Paul Balanca

    cs.LGarXiv:2609.02846v12026
  14. Aesthetic-based Clothing Recommendation

    Wenhui Yu, Huidi Zhang, Xiangnan He +3

    cs.IRcs.LGstat.MLarXiv:1809.05822v12018
  15. Recursive Introspection: Teaching Language Model Agents How to Self-Improve

    Yuxiao Qu, Tianjun Zhang, Naman Garg +1

    cs.LGcs.AIcs.CLarXiv:2407.18219v22024
  16. GRADSOLVE: fast exact gradients for ODE ensembles on GPUs

    Alessio Spurio Mancini

    cs.MScs.DCcs.LGarXiv:2609.02876v12026
  17. Learning Independent Causal Mechanisms

    Giambattista Parascandolo, Niki Kilbertus, Mateo Rojas-Carulla +1

    cs.LGstat.MLarXiv:1712.00961v52017
  18. Real-Time Radio Technology and Modulation Classification via an LSTM Auto-Encoder

    Ziqi Ke, Haris Vikalo

    eess.SPcs.LGarXiv:2011.08295v12020
  19. The Early Phase of Neural Network Training

    Jonathan Frankle, David J. Schwab, Ari S. Morcos

    cs.LGcs.NEstat.MLarXiv:2002.10365v12020
  20. Minimax Weight and Q-Function Learning for Off-Policy Evaluation

    Masatoshi Uehara, Jiawei Huang, Nan Jiang

    cs.LGstat.MLarXiv:1910.12809v42019
  21. Sub-graph Contrast for Scalable Self-Supervised Graph Representation Learning

    Yizhu Jiao, Yun Xiong, Jiawei Zhang +3

    cs.LGstat.MLarXiv:2009.10273v32020
  22. Requirements Engineering for Machine Learning: Perspectives from Data Scientists

    Andreas Vogelsang, Markus Borg

    cs.LGcs.SEarXiv:1908.04674v12019
  23. The Break-Even Point on Optimization Trajectories of Deep Neural Networks

    Stanislaw Jastrzebski, Maciej Szymczak, Stanislav Fort +4

    cs.LGstat.MLarXiv:2002.09572v12020
  24. TaRA: Training-Aware Low-Rank Adaptation Initialization

    Taehyeon Kim, Eunhyeok Park

    cs.CLcs.AIcs.LGarXiv:2609.02639v12026
  25. HDMI: High-order Deep Multiplex Infomax

    Baoyu Jing, Chanyoung Park, Hanghang Tong

    cs.LGcs.ITcs.SIarXiv:2102.07810v52021
  26. HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding

    Zhaorun Chen, Zhuokai Zhao, Hongyin Luo +3

    cs.CVcs.AIcs.LGarXiv:2403.00425v22024
  27. Sampling Can Be Faster Than Optimization

    Yi-An Ma, Yuansi Chen, Chi Jin +2

    stat.MLcs.LGarXiv:1811.08413v22018
  28. Graph Information Bottleneck for Subgraph Recognition

    Junchi Yu, Tingyang Xu, Yu Rong +3

    cs.LGstat.MLarXiv:2010.05563v12020
  29. Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers

    Giovanni Bonetta, Matteo Merler, Davide Zago +2

    cs.AIcs.CLcs.LGarXiv:2609.01567v22026
  30. Towards Understanding Sharpness-Aware Minimization

    Maksym Andriushchenko, Nicolas Flammarion

    cs.LGarXiv:2206.06232v12022
  31. FairFace: Face Attribute Dataset for Balanced Race, Gender, and Age

    Kimmo Kärkkäinen, Jungseock Joo

    cs.CVcs.LGarXiv:1908.04913v12019
  32. Artificial Intelligence and Life in 2030: The One Hundred Year Study on Artificial Intelligence

    Peter Stone, Rodney Brooks, Erik Brynjolfsson +14

    cs.CYcs.AIcs.LGarXiv:2211.06318v12022
  33. UKP-Athene: Multi-Sentence Textual Entailment for Claim Verification

    Andreas Hanselowski, Hao Zhang, Zile Li +4

    cs.IRcs.AIcs.CLarXiv:1809.01479v52018
  34. Socially Compliant Navigation through Raw Depth Inputs with Generative Adversarial Imitation Learning

    Lei Tai, Jingwei Zhang, Ming Liu +1

    cs.ROcs.AIcs.LGarXiv:1710.02543v22017
  35. A CHAID Based Performance Prediction Model in Educational Data Mining

    M. Ramaswami, R. Bhaskaran

    cs.LGarXiv:1002.1144v12010
  36. Ultimate tensorization: compressing convolutional and FC layers alike

    Timur Garipov, Dmitry Podoprikhin, Alexander Novikov +1

    cs.LGarXiv:1611.03214v12016
  37. SciRepEval: A Multi-Format Benchmark for Scientific Document Representations

    Amanpreet Singh, Mike D'Arcy, Arman Cohan +2

    cs.CLcs.AIcs.IRarXiv:2211.13308v42022
  38. Counterfactual Off-Policy Evaluation with Gumbel-Max Structural Causal Models

    Michael Oberst, David Sontag

    cs.LGstat.MLarXiv:1905.05824v32019
  39. Entropy and mutual information in models of deep neural networks

    Marylou Gabrié, Andre Manoel, Clément Luneau +4

    cs.LGcond-mat.dis-nncs.ITarXiv:1805.09785v22018
  40. Partial Is Better Than All: Revisiting Fine-tuning Strategy for Few-shot Learning

    Zhiqiang Shen, Zechun Liu, Jie Qin +2

    cs.CVcs.AIcs.LGarXiv:2102.03983v12021
  41. Efficiently Estimating Optimal Hyperparameter Scaling Laws through Power-Law Entropy Search

    Zhiliang Chen, Sebastian Ament, David Eriksson +4

    cs.LGcs.AIarXiv:2609.01431v22026
  42. Big Data Meet Cyber-Physical Systems: A Panoramic Survey

    Rachad Atat, Lingjia Liu, Jinsong Wu +3

    cs.LGstat.MLarXiv:1810.12399v12018
  43. Deep Learning for Human Affect Recognition: Insights and New Developments

    Philipp V. Rouast, Marc T. P. Adam, Raymond Chiong

    cs.LGcs.AIcs.CVarXiv:1901.02884v12019
  44. A Survey of Deep Learning for Mathematical Reasoning

    Pan Lu, Liang Qiu, Wenhao Yu +2

    cs.AIcs.CLcs.CVarXiv:2212.10535v22022
  45. Cognitive Psychology for Deep Neural Networks: A Shape Bias Case Study

    Samuel Ritter, David G. T. Barrett, Adam Santoro +1

    stat.MLcs.CVcs.LGarXiv:1706.08606v22017
  46. Could a Large Language Model be Conscious?

    David J. Chalmers

    cs.AIcs.CLcs.LGarXiv:2303.07103v32023
  47. Loss Aware Post-training Quantization

    Yury Nahshan, Brian Chmiel, Chaim Baskin +4

    cs.LGcs.CVarXiv:1911.07190v22019
  48. Learning Neural Templates for Text Generation

    Sam Wiseman, Stuart M. Shieber, Alexander M. Rush

    cs.CLcs.LGarXiv:1808.10122v32018
  49. One-pass Multi-task Networks with Cross-task Guided Attention for Brain Tumor Segmentation

    Chenhong Zhou, Changxing Ding, Xinchao Wang +2

    cs.CVcs.AIcs.LGarXiv:1906.01796v22019
  50. An Empirical Study on Robustness to Spurious Correlations using Pre-trained Language Models

    Lifu Tu, Garima Lalwani, Spandana Gella +1

    cs.CLcs.LGarXiv:2007.06778v32020
  51. Counterfactual Memorization in Neural Language Models

    Chiyuan Zhang, Daphne Ippolito, Katherine Lee +3

    cs.CLcs.AIcs.LGarXiv:2112.12938v22021
  52. A Simple Exponential Family Framework for Zero-Shot Learning

    Vinay Kumar Verma, Piyush Rai

    cs.LGcs.CVstat.MLarXiv:1707.08040v32017
  53. Embodied Question Answering in Photorealistic Environments with Point Cloud Perception

    Erik Wijmans, Samyak Datta, Oleksandr Maksymets +6

    cs.CVcs.AIcs.CLarXiv:1904.03461v12019
  54. The Ingredients of Real-World Robotic Reinforcement Learning

    Henry Zhu, Justin Yu, Abhishek Gupta +5

    cs.LGcs.ROstat.MLarXiv:2004.12570v12020
  55. Pushing Large Language Models to the 6G Edge: Vision, Challenges, and Opportunities

    Zheng Lin, Guanqiao Qu, Qiyuan Chen +3

    cs.LGcs.AIarXiv:2309.16739v42023
  56. Efficient Guided Generation for Large Language Models

    Brandon T. Willard, Rémi Louf

    cs.CLcs.LGarXiv:2307.09702v42023
  57. Hearing the Whispers: Black-Box Membership Inference Attacks on Finetuned TTS Models

    Kunlin Cai, Kaiyuan Zhang, Zihang Xiang +4

    cs.CRcs.LGcs.SDarXiv:2609.01723v12026
  58. Multi-Agent Reinforcement Learning for Active Voltage Control on Power Distribution Networks

    Jianhong Wang, Wangkun Xu, Yunjie Gu +2

    cs.LGcs.MAarXiv:2110.14300v52021
  59. Don't forget, there is more than forgetting: new metrics for Continual Learning

    Natalia Díaz-Rodríguez, Vincenzo Lomonaco, David Filliat +1

    cs.AIcs.CVcs.LGarXiv:1810.13166v12018
  60. FMore: An Incentive Scheme of Multi-dimensional Auction for Federated Learning in MEC

    Rongfei Zeng, Shixun Zhang, Jiaqi Wang +1

    cs.LGcs.GTstat.MLarXiv:2002.09699v12020