Machine Learning

Papers filed under cs.LG on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

9,961 to 10,020 of 19,974

  1. Self-concordant analysis for logistic regression

    Francis Bach

    cs.LGmath.STarXiv:0910.4627v12009
  2. What Large Language Models Know and What People Think They Know

    Mark Steyvers, Heliodoro Tejeda, Aakriti Kumar +5

    cs.LGcs.AIcs.CLarXiv:2401.13835v22024
  3. Delving into the Devils of Bird's-eye-view Perception: A Review, Evaluation and Recipe

    Hongyang Li, Chonghao Sima, Jifeng Dai +19

    cs.CVcs.LGcs.ROarXiv:2209.05324v42022
  4. To Cluster, or Not to Cluster: An Analysis of Clusterability Methods

    A. Adolfsson, M. Ackerman, N. C. Brownstein

    stat.MLcs.LGarXiv:1808.08317v12018
  5. Deep Learning for Anomaly Detection in Log Data: A Survey

    Max Landauer, Sebastian Onder, Florian Skopik +1

    cs.LGarXiv:2207.03820v22022
  6. ST-GAN: Spatial Transformer Generative Adversarial Networks for Image Compositing

    Chen-Hsuan Lin, Ersin Yumer, Oliver Wang +2

    cs.CVcs.LGarXiv:1803.01837v12018
  7. Synthesizing Programs for Images using Reinforced Adversarial Learning

    Yaroslav Ganin, Tejas Kulkarni, Igor Babuschkin +2

    cs.CVcs.LGstat.MLarXiv:1804.01118v12018
  8. Code Prediction by Feeding Trees to Transformers

    Seohyun Kim, Jinman Zhao, Yuchi Tian +1

    cs.SEcs.LGarXiv:2003.13848v42020
  9. The PyTorch-Kaldi Speech Recognition Toolkit

    Mirco Ravanelli, Titouan Parcollet, Yoshua Bengio

    eess.AScs.CLcs.LGarXiv:1811.07453v22018
  10. Verifying Properties of Binarized Deep Neural Networks

    Nina Narodytska, Shiva Prasad Kasiviswanathan, Leonid Ryzhyk +2

    stat.MLcs.AIcs.CRarXiv:1709.06662v22017
  11. Code Generation as a Dual Task of Code Summarization

    Bolin Wei, Ge Li, Xin Xia +2

    cs.LGcs.AIcs.SEarXiv:1910.05923v12019
  12. Towards Understanding Generalization of Deep Learning: Perspective of Loss Landscapes

    Lei Wu, Zhanxing Zhu, Weinan E

    cs.LGcs.AIstat.MLarXiv:1706.10239v22017
  13. Linguistic Distance Segregates Latent Representations in Automatic Speech Recognition Systems

    Ting-Hui Cheng, Line Katrine Harder Clemmensen, Sneha Das

    cs.CLcs.LGarXiv:2608.30853v12026
  14. SPViT: Enabling Faster Vision Transformers via Soft Token Pruning

    Zhenglun Kong, Peiyan Dong, Xiaolong Ma +9

    cs.CVcs.AIcs.ARarXiv:2112.13890v22021
  15. The Unsurprising Effectiveness of Pre-Trained Vision Models for Control

    Simone Parisi, Aravind Rajeswaran, Senthil Purushwalkam +1

    cs.CVcs.AIcs.LGarXiv:2203.03580v22022
  16. Deep Learning Enables Automatic Detection and Segmentation of Brain Metastases on Multi-Sequence MRI

    Endre Grøvik, Darvin Yi, Michael Iv +3

    eess.IVcs.LGstat.MLarXiv:1903.07988v12019
  17. Anomal-E: A Self-Supervised Network Intrusion Detection System based on Graph Neural Networks

    Evan Caville, Wai Weng Lo, Siamak Layeghy +1

    cs.LGcs.AIcs.CRarXiv:2207.06819v52022
  18. TopoCompress: Long Context Compression via Graph-Wired Semantic Trajectories

    Daniel Agyei Asante, Yang Li

    cs.CLcs.LGarXiv:2608.30811v12026
  19. Efficient and Scalable Bayesian Neural Nets with Rank-1 Factors

    Michael W. Dusenberry, Ghassen Jerfel, Yeming Wen +5

    cs.LGstat.MLarXiv:2005.07186v22020
  20. Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models

    Asma Ghandeharioun, Avi Caciularu, Adam Pearce +2

    cs.CLcs.AIcs.LGarXiv:2401.06102v42024
  21. Inferring clonal evolution of tumors from single nucleotide somatic mutations

    Wei Jiao, Shankar Vembu, Amit G. Deshwar +2

    cs.LGq-bio.PEq-bio.QMarXiv:1210.3384v42012
  22. Foundation Models for Generalist Geospatial Artificial Intelligence

    Johannes Jakubik, Sujit Roy, C. E. Phillips +30

    cs.CVcs.LGarXiv:2310.18660v22023
  23. Simultaneous Navigation and Radio Mapping for Cellular-Connected UAV with Deep Reinforcement Learning

    Yong Zeng, Xiaoli Xu, Shi Jin +1

    eess.SPcs.ITcs.LGarXiv:2003.07574v12020
  24. Deep Learning for Spacecraft Pose Estimation from Photorealistic Rendering

    Pedro F. Proenca, Yang Gao

    cs.CVcs.LGcs.ROarXiv:1907.04298v22019
  25. When Deep Learning Met Code Search

    Jose Cambronero, Hongyu Li, Seohyun Kim +2

    cs.SEcs.CLcs.LGarXiv:1905.03813v42019
  26. A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix Groups

    Marc Finzi, Max Welling, Andrew Gordon Wilson

    cs.LGmath.DSstat.MLarXiv:2104.09459v12021
  27. No bad local minima: Data independent training error guarantees for multilayer neural networks

    Daniel Soudry, Yair Carmon

    stat.MLcs.LGcs.NEarXiv:1605.08361v22016
  28. Neurotoxin: Durable Backdoors in Federated Learning

    Zhengming Zhang, Ashwinee Panda, Linyue Song +5

    cs.CRcs.AIcs.LGarXiv:2206.10341v12022
  29. $β$-Variational autoencoders and transformers for reduced-order modelling of fluid flows

    Alberto Solera-Rico, Carlos Sanmiguel Vila, M. A. Gómez +4

    physics.flu-dyncs.LGarXiv:2304.03571v22023
  30. Active Deep Learning for Classification of Hyperspectral Images

    Peng Liu, Hui Zhang, Kie B. Eom

    cs.LGcs.CVstat.MLarXiv:1611.10031v12016
  31. FastDiff: A Fast Conditional Diffusion Model for High-Quality Speech Synthesis

    Rongjie Huang, Max W. Y. Lam, Jun Wang +4

    eess.AScs.LGcs.SDarXiv:2204.09934v12022
  32. Iterative Views Agreement: An Iterative Low-Rank based Structured Optimization Method to Multi-View Spectral Clustering

    Yang Wang, Wenjie Zhang, Lin Wu +3

    cs.LGstat.MLarXiv:1608.05560v12016
  33. Learning Fast Samplers for Diffusion Models by Differentiating Through Sample Quality

    Daniel Watson, William Chan, Jonathan Ho +1

    cs.LGarXiv:2202.05830v12022
  34. PromptFL: Let Federated Participants Cooperatively Learn Prompts Instead of Models -- Federated Learning in Age of Foundation Model

    Tao Guo, Song Guo, Junxiao Wang +1

    cs.LGarXiv:2208.11625v12022
  35. GCNet: Graph Completion Network for Incomplete Multimodal Learning in Conversation

    Zheng Lian, Lan Chen, Licai Sun +2

    cs.LGcs.CLarXiv:2203.02177v22022
  36. Supervised Contrastive Replay: Revisiting the Nearest Class Mean Classifier in Online Class-Incremental Continual Learning

    Zheda Mai, Ruiwen Li, Hyunwoo Kim +1

    cs.LGcs.AIcs.CVarXiv:2103.13885v32021
  37. Self-Supervised Models are Continual Learners

    Enrico Fini, Victor G. Turrisi da Costa, Xavier Alameda-Pineda +3

    cs.CVcs.LGarXiv:2112.04215v22021
  38. Learning Near Optimal Policies with Low Inherent Bellman Error

    Andrea Zanette, Alessandro Lazaric, Mykel Kochenderfer +1

    cs.LGcs.AIarXiv:2003.00153v32020
  39. GEAR: Graph-based Evidence Aggregating and Reasoning for Fact Verification

    Jie Zhou, Xu Han, Cheng Yang +4

    cs.CLcs.AIcs.LGarXiv:1908.01843v12019
  40. A Survey on Incomplete Multi-view Clustering

    Jie Wen, Zheng Zhang, Lunke Fei +4

    cs.LGcs.MMarXiv:2208.08040v12022
  41. An Agentic Retrobiosynthesis Framework with Learned Frontier Selection

    Philippe Meyer, Guillaume Gricourt, Thomas Duigou +2

    cs.CLcs.AIcs.LGarXiv:2608.30702v12026
  42. How to Fool Radiologists with Generative Adversarial Networks? A Visual Turing Test for Lung Cancer Diagnosis

    Maria J. M. Chuquicusma, Sarfaraz Hussein, Jeremy Burt +1

    cs.CVcs.AIcs.LGarXiv:1710.09762v22017
  43. Type-Constrained Representation Learning in Knowledge Graphs

    Denis Krompaß, Stephan Baier, Volker Tresp

    cs.AIcs.LGarXiv:1508.02593v22015
  44. NeVAE: A Deep Generative Model for Molecular Graphs

    Bidisha Samanta, Abir De, Gourhari Jana +3

    cs.LGphysics.soc-phstat.MLarXiv:1802.05283v42018
  45. Extreme Compression of Large Language Models via Additive Quantization

    Vage Egiazarian, Andrei Panferov, Denis Kuznedelev +3

    cs.LGcs.CLarXiv:2401.06118v42024
  46. Benchmarking Deep Learning Interpretability in Time Series Predictions

    Aya Abdelsalam Ismail, Mohamed Gunady, Héctor Corrada Bravo +1

    cs.LGstat.MLarXiv:2010.13924v12020
  47. BiG-SURE - Bipartite Graph for Semantic Uncertainty and Reliability Estimation of LLMs

    Debarpan Bhattacharya, Malay Phadke, Sriram Ganapathy

    cs.CLcs.AIcs.LGarXiv:2608.30646v12026
  48. What It Costs to Compose, Rebuild, and Correct Precomputed Memory

    Asa Shepard

    cs.CLcs.LGarXiv:2608.30647v12026
  49. GMTS: Gradient Magnitude-based Token Selection Improves RLVR Training for LLM Reasoning

    Outongyi Lv, Yuanwei Zhang, Xiaoqun Zhang

    cs.CLcs.AIcs.LGarXiv:2608.30632v12026
  50. Continuous State-Space Models for Optimal Sepsis Treatment - a Deep Reinforcement Learning Approach

    Aniruddh Raghu, Matthieu Komorowski, Leo Anthony Celi +2

    cs.LGarXiv:1705.08422v12017
  51. UniTime: A Language-Empowered Unified Model for Cross-Domain Time Series Forecasting

    Xu Liu, Junfeng Hu, Yuan Li +4

    cs.LGarXiv:2310.09751v32023
  52. Pushing the Boundaries of Boundary Detection using Deep Learning

    Iasonas Kokkinos

    cs.CVcs.LGarXiv:1511.07386v22015
  53. A Machine-Learning Approach for Earthquake Magnitude Estimation

    S. Mostafa Mousavi, Gregory C. Beroza

    physics.geo-phcs.AIcs.LGarXiv:1911.05975v12019
  54. Reading the News: Adapting Large Language Models to Swedish Journalism Through Continued Pre-Training

    Lukas Borggren, Jenny Kunz, Marco Kuhlmann

    cs.CLcs.AIcs.LGarXiv:2608.30609v12026
  55. Scaling Robot Learning with Semantically Imagined Experience

    Tianhe Yu, Ted Xiao, Austin Stone +10

    cs.ROcs.AIcs.CLarXiv:2302.11550v12023
  56. Foundational Challenges in Assuring Alignment and Safety of Large Language Models

    Usman Anwar, Abulhair Saparov, Javier Rando +39

    cs.LGcs.AIcs.CLarXiv:2404.09932v22024
  57. ColO-RAN: Developing Machine Learning-based xApps for Open RAN Closed-loop Control on Programmable Experimental Platforms

    Michele Polese, Leonardo Bonati, Salvatore D'Oro +2

    cs.NIcs.LGarXiv:2112.09559v22021
  58. Generating Classification Weights with GNN Denoising Autoencoders for Few-Shot Learning

    Spyros Gidaris, Nikos Komodakis

    cs.CVcs.LGarXiv:1905.01102v12019
  59. Delta Tuning: A Comprehensive Study of Parameter Efficient Methods for Pre-trained Language Models

    Ning Ding, Yujia Qin, Guang Yang +17

    cs.CLcs.AIcs.LGarXiv:2203.06904v22022
  60. A Primer on PAC-Bayesian Learning

    Benjamin Guedj

    stat.MLcs.LGarXiv:1901.05353v32019