Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

8,821 to 8,880 of 20,454

  1. Deep Learning for Human Affect Recognition: Insights and New Developments

    Philipp V. Rouast, Marc T. P. Adam, Raymond Chiong

    cs.LGcs.AIcs.CVarXiv:1901.02884v12019
  2. A Survey of Deep Learning for Mathematical Reasoning

    Pan Lu, Liang Qiu, Wenhao Yu +2

    cs.AIcs.CLcs.CVarXiv:2212.10535v22022
  3. Cognitive Psychology for Deep Neural Networks: A Shape Bias Case Study

    Samuel Ritter, David G. T. Barrett, Adam Santoro +1

    stat.MLcs.CVcs.LGarXiv:1706.08606v22017
  4. Could a Large Language Model be Conscious?

    David J. Chalmers

    cs.AIcs.CLcs.LGarXiv:2303.07103v32023
  5. Loss Aware Post-training Quantization

    Yury Nahshan, Brian Chmiel, Chaim Baskin +4

    cs.LGcs.CVarXiv:1911.07190v22019
  6. Learning Neural Templates for Text Generation

    Sam Wiseman, Stuart M. Shieber, Alexander M. Rush

    cs.CLcs.LGarXiv:1808.10122v32018
  7. One-pass Multi-task Networks with Cross-task Guided Attention for Brain Tumor Segmentation

    Chenhong Zhou, Changxing Ding, Xinchao Wang +2

    cs.CVcs.AIcs.LGarXiv:1906.01796v22019
  8. An Empirical Study on Robustness to Spurious Correlations using Pre-trained Language Models

    Lifu Tu, Garima Lalwani, Spandana Gella +1

    cs.CLcs.LGarXiv:2007.06778v32020
  9. Counterfactual Memorization in Neural Language Models

    Chiyuan Zhang, Daphne Ippolito, Katherine Lee +3

    cs.CLcs.AIcs.LGarXiv:2112.12938v22021
  10. A Simple Exponential Family Framework for Zero-Shot Learning

    Vinay Kumar Verma, Piyush Rai

    cs.LGcs.CVstat.MLarXiv:1707.08040v32017
  11. Embodied Question Answering in Photorealistic Environments with Point Cloud Perception

    Erik Wijmans, Samyak Datta, Oleksandr Maksymets +6

    cs.CVcs.AIcs.CLarXiv:1904.03461v12019
  12. The Ingredients of Real-World Robotic Reinforcement Learning

    Henry Zhu, Justin Yu, Abhishek Gupta +5

    cs.LGcs.ROstat.MLarXiv:2004.12570v12020
  13. Pushing Large Language Models to the 6G Edge: Vision, Challenges, and Opportunities

    Zheng Lin, Guanqiao Qu, Qiyuan Chen +3

    cs.LGcs.AIarXiv:2309.16739v42023
  14. Efficient Guided Generation for Large Language Models

    Brandon T. Willard, Rémi Louf

    cs.CLcs.LGarXiv:2307.09702v42023
  15. Hearing the Whispers: Black-Box Membership Inference Attacks on Finetuned TTS Models

    Kunlin Cai, Kaiyuan Zhang, Zihang Xiang +4

    cs.CRcs.LGcs.SDarXiv:2609.01723v12026
  16. Multi-Agent Reinforcement Learning for Active Voltage Control on Power Distribution Networks

    Jianhong Wang, Wangkun Xu, Yunjie Gu +2

    cs.LGcs.MAarXiv:2110.14300v52021
  17. Don't forget, there is more than forgetting: new metrics for Continual Learning

    Natalia Díaz-Rodríguez, Vincenzo Lomonaco, David Filliat +1

    cs.AIcs.CVcs.LGarXiv:1810.13166v12018
  18. FMore: An Incentive Scheme of Multi-dimensional Auction for Federated Learning in MEC

    Rongfei Zeng, Shixun Zhang, Jiaqi Wang +1

    cs.LGcs.GTstat.MLarXiv:2002.09699v12020
  19. Hurtful Words: Quantifying Biases in Clinical Contextual Word Embeddings

    Haoran Zhang, Amy X. Lu, Mohamed Abdalla +2

    cs.CLcs.CYcs.LGarXiv:2003.11515v12020
  20. Generative Image Modeling Using Spatial LSTMs

    Lucas Theis, Matthias Bethge

    stat.MLcs.CVcs.LGarXiv:1506.03478v22015
  21. Distributed Learning with Compressed Gradient Differences

    Konstantin Mishchenko, Eduard Gorbunov, Martin Takáč +1

    cs.LGmath.OCstat.MLarXiv:1901.09269v32019
  22. Model-based Reinforcement Learning and the Eluder Dimension

    Ian Osband, Benjamin Van Roy

    stat.MLcs.LGarXiv:1406.1853v22014
  23. ReGVD: Revisiting Graph Neural Networks for Vulnerability Detection

    Van-Anh Nguyen, Dai Quoc Nguyen, Van Nguyen +3

    cs.LGcs.CRarXiv:2110.07317v32021
  24. Unsupervised Multi-Task Feature Learning on Point Clouds

    Kaveh Hassani, Mike Haley

    cs.CVcs.LGarXiv:1910.08207v12019
  25. Manifold-Aware General Coded Computing for Straggler-Resilient Distributed Computing

    Parsa Moradi, Mohammad Ali Maddah-Ali

    cs.LGcs.ITarXiv:2609.00552v12026
  26. OR-Transformer: Scaling Real-Time Decision-Making to 1,000 Items

    Shuze Daniel Liu, David Simchi-Levi, Claire Chen +2

    cs.LGarXiv:2609.01933v12026
  27. Interpretable Adversarial Perturbation in Input Embedding Space for Text

    Motoki Sato, Jun Suzuki, Hiroyuki Shindo +1

    cs.LGcs.CLstat.MLarXiv:1805.02917v12018
  28. DRAMA: Joint Risk Localization and Captioning in Driving

    Srikanth Malla, Chiho Choi, Isht Dwivedi +2

    cs.CVcs.AIcs.LGarXiv:2209.10767v22022
  29. Non-Local Graph Neural Networks

    Meng Liu, Zhengyang Wang, Shuiwang Ji

    cs.LGstat.MLarXiv:2005.14612v22020
  30. A Unified View on Graph Neural Networks as Graph Signal Denoising

    Yao Ma, Xiaorui Liu, Tong Zhao +3

    cs.LGstat.MLarXiv:2010.01777v22020
  31. Data-dependent Initializations of Convolutional Neural Networks

    Philipp Krähenbühl, Carl Doersch, Jeff Donahue +1

    cs.CVcs.LGarXiv:1511.06856v32015
  32. MPAF: Model Poisoning Attacks to Federated Learning based on Fake Clients

    Xiaoyu Cao, Neil Zhenqiang Gong

    cs.CRcs.LGarXiv:2203.08669v22022
  33. Exact Risk-Complexity Laws for Projective Boundaries in Scenario Optimization and Distribution-Free Certification

    Giuseppe C. Calafiore

    eess.SYcs.LGmath.OCarXiv:2609.01355v12026
  34. A Mathematical Theory of Reusable Neural Bases for Network Compression

    Binshuai Wang, Peng Wei

    cs.LGcs.AIarXiv:2609.01550v22026
  35. FCN-Transformer Feature Fusion for Polyp Segmentation

    Edward Sanderson, Bogdan J. Matuszewski

    eess.IVcs.CVcs.LGarXiv:2208.08352v12022
  36. Self-paced Ensemble for Highly Imbalanced Massive Data Classification

    Zhining Liu, Wei Cao, Zhifeng Gao +4

    cs.LGcs.AIstat.MLarXiv:1909.03500v32019
  37. Panoptic Lifting for 3D Scene Understanding with Neural Fields

    Yawar Siddiqui, Lorenzo Porzi, Samuel Rota Buló +4

    cs.CVcs.LGarXiv:2212.09802v12022
  38. Unbalanced minibatch Optimal Transport; applications to Domain Adaptation

    Kilian Fatras, Thibault Séjourné, Nicolas Courty +1

    cs.LGmath.STstat.MLarXiv:2103.03606v12021
  39. Not Just Privacy: Improving Performance of Private Deep Learning in Mobile Cloud

    Ji Wang, Jianguo Zhang, Weidong Bao +3

    cs.LGcs.AIcs.DCarXiv:1809.03428v32018
  40. Optimal Schemes for Discrete Distribution Estimation under Locally Differential Privacy

    Min Ye, Alexander Barg

    cs.LGcs.ITarXiv:1702.00610v12017
  41. CHASE-SQL: Multi-Path Reasoning and Preference Optimized Candidate Selection in Text-to-SQL

    Mohammadreza Pourreza, Hailong Li, Ruoxi Sun +7

    cs.LGcs.AIcs.CLarXiv:2410.01943v12024
  42. PaperQA: Retrieval-Augmented Generative Agent for Scientific Research

    Jakub Lála, Odhran O'Donoghue, Aleksandar Shtedritski +3

    cs.CLcs.AIcs.LGarXiv:2312.07559v22023
  43. An Intuitive Tutorial to Gaussian Process Regression

    Jie Wang

    stat.MLcs.LGcs.ROarXiv:2009.10862v52020
  44. Reducing hallucination in structured outputs via Retrieval-Augmented Generation

    Patrice Béchard, Orlando Marquez Ayala

    cs.LGcs.AIcs.CLarXiv:2404.08189v12024
  45. SMACv2: An Improved Benchmark for Cooperative Multi-Agent Reinforcement Learning

    Benjamin Ellis, Jonathan Cook, Skander Moalla +5

    cs.LGcs.MAarXiv:2212.07489v22022
  46. Atomic Convolutional Networks for Predicting Protein-Ligand Binding Affinity

    Joseph Gomes, Bharath Ramsundar, Evan N. Feinberg +1

    cs.LGphysics.chem-phstat.MLarXiv:1703.10603v12017
  47. Privacy Amplification by Iteration

    Vitaly Feldman, Ilya Mironov, Kunal Talwar +1

    cs.LGcs.CRcs.DSarXiv:1808.06651v22018
  48. Few-shot Slot Tagging with Collapsed Dependency Transfer and Label-enhanced Task-adaptive Projection Network

    Yutai Hou, Wanxiang Che, Yongkui Lai +4

    cs.CLcs.LGarXiv:2006.05702v12020
  49. Deep learning approach based on dimensionality reduction for designing electromagnetic nanostructures

    Yashar Kiarashinejad, Sajjad Abdollahramezani, Ali Adibi

    cs.LGphysics.app-phstat.MLarXiv:1902.03865v32019
  50. Deep Learning for Post-Processing Ensemble Weather Forecasts

    Peter Grönquist, Chengyuan Yao, Tal Ben-Nun +4

    cs.LGeess.SPphysics.ao-pharXiv:2005.08748v22020
  51. Deep Learning for Launching and Mitigating Wireless Jamming Attacks

    Tugba Erpek, Yalin E. Sagduyu, Yi Shi

    cs.NIcs.LGstat.MLarXiv:1807.02567v22018
  52. Learning to Solve NP-Complete Problems - A Graph Neural Network for Decision TSP

    Marcelo O. R. Prates, Pedro H. C. Avelar, Henrique Lemos +2

    cs.LGcs.AIcs.NEarXiv:1809.02721v32018
  53. Just How Toxic is Data Poisoning? A Unified Benchmark for Backdoor and Data Poisoning Attacks

    Avi Schwarzschild, Micah Goldblum, Arjun Gupta +2

    cs.LGcs.CRcs.CVarXiv:2006.12557v32020
  54. Continuous Graph Neural Networks

    Louis-Pascal A. C. Xhonneux, Meng Qu, Jian Tang

    cs.LGstat.MLarXiv:1912.00967v32019
  55. A Deep Value-network Based Approach for Multi-Driver Order Dispatching

    Xiaocheng Tang, Zhiwei Qin, Fan Zhang +5

    cs.LGcs.AIarXiv:2106.04493v12021
  56. SemanticAdv: Generating Adversarial Examples via Attribute-conditional Image Editing

    Haonan Qiu, Chaowei Xiao, Lei Yang +3

    cs.LGcs.CRcs.CVarXiv:1906.07927v42019
  57. Streaming automatic speech recognition with the transformer model

    Niko Moritz, Takaaki Hori, Jonathan Le Roux

    cs.SDcs.CLcs.LGarXiv:2001.02674v52020
  58. Incorporating Symmetry into Deep Dynamics Models for Improved Generalization

    Rui Wang, Robin Walters, Rose Yu

    cs.LGmath.RTstat.MLarXiv:2002.03061v42020
  59. Emformer: Efficient Memory Transformer Based Acoustic Model For Low Latency Streaming Speech Recognition

    Yangyang Shi, Yongqiang Wang, Chunyang Wu +5

    cs.SDcs.CLcs.LGarXiv:2010.10759v42020
  60. LatentPress: Context Compression Beyond Text and Vision

    Zhengze Zhou, Hejian Sang

    cs.LGcs.AIarXiv:2609.01507v22026