Every paper with a summary

Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

55,441 to 55,500 of 61,162

  1. AnalogRetriever: Learning Cross-Modal Representations for Analog Circuit Retrieval

    Yihan Wang, Lei Li, Yao Lai +2

    cs.CVcs.AIarXiv:2604.23195v12026
  2. DeeperCut: A Deeper, Stronger, and Faster Multi-Person Pose Estimation Model

    Eldar Insafutdinov, Leonid Pishchulin, Bjoern Andres +2

    cs.CVarXiv:1605.03170v32016
  3. OceanPile: A Large-Scale Multimodal Ocean Corpus for Foundation Models

    Yida Xue, Ningyu Zhang, Tingwei Wu +5

    cs.MMcs.AIcs.CLarXiv:2605.00877v22026
  4. Data Augmentation Generative Adversarial Networks

    Antreas Antoniou, Amos Storkey, Harrison Edwards

    stat.MLcs.CVcs.LGarXiv:1711.04340v32017
  5. Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentation

    Simone Mosco, Daniel Fusaro, Alberto Pretto

    cs.CVcs.ROarXiv:2604.23604v12026
  6. Wasserstein Auto-Encoders

    Ilya Tolstikhin, Olivier Bousquet, Sylvain Gelly +1

    stat.MLcs.LGarXiv:1711.01558v42017
  7. AgentBench: Evaluating LLMs as Agents

    Xiao Liu, Hao Yu, Hanchen Zhang +19

    cs.AIcs.CLcs.LGarXiv:2308.03688v32023
  8. Quantum Kernel Advantage over Classical Collapse in Medical Foundation Model Embeddings

    Sebastian Cajas Ordóñez, Felipe Ocampo Osorio, Dax Enshan Koh +10

    quant-phcs.AIarXiv:2604.24597v12026
  9. VIBE: Video Inference for Human Body Pose and Shape Estimation

    Muhammed Kocabas, Nikos Athanasiou, Michael J. Black

    cs.CVarXiv:1912.05656v32019
  10. DenseFusion: 6D Object Pose Estimation by Iterative Dense Fusion

    Chen Wang, Danfei Xu, Yuke Zhu +4

    cs.CVcs.ROarXiv:1901.04780v12019
  11. Learning Spatio-Temporal Transformer for Visual Tracking

    Bin Yan, Houwen Peng, Jianlong Fu +2

    cs.CVarXiv:2103.17154v12021
  12. AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark

    Hongxin Li, Xiping Wang, Jingran Su +4

    cs.CVarXiv:2604.24441v12026
  13. Secure Transmission with Multiple Antennas: The MISOME Wiretap Channel

    Ashish Khisti, Gregory Wornell

    cs.ITarXiv:0708.4219v12007
  14. CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval

    Huaishao Luo, Lei Ji, Ming Zhong +4

    cs.CVarXiv:2104.08860v22021
  15. GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction

    Hongxin Li, Yuntao Chen, Zhaoxiang Zhang

    cs.CVarXiv:2604.23941v12026
  16. S^3-Rec: Self-Supervised Learning for Sequential Recommendation with Mutual Information Maximization

    Kun Zhou, Hui Wang, Wayne Xin Zhao +5

    cs.IRcs.LGarXiv:2008.07873v12020
  17. StackGAN++: Realistic Image Synthesis with Stacked Generative Adversarial Networks

    Han Zhang, Tao Xu, Hongsheng Li +4

    cs.CVcs.AIstat.MLarXiv:1710.10916v32017
  18. Composition-based Multi-Relational Graph Convolutional Networks

    Shikhar Vashishth, Soumya Sanyal, Vikram Nitin +1

    cs.LGstat.MLarXiv:1911.03082v22019
  19. ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

    Lin Chen, Jinsong Li, Xiaoyi Dong +5

    cs.CVarXiv:2311.12793v22023
  20. Enhanced LSTM for Natural Language Inference

    Qian Chen, Xiaodan Zhu, Zhenhua Ling +3

    cs.CLarXiv:1609.06038v32016
  21. BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate

    Arnon Mazza, Elad Levi

    cs.CLcs.AIcs.LGarXiv:2604.25203v12026
  22. Symmetric Cross Entropy for Robust Learning with Noisy Labels

    Yisen Wang, Xingjun Ma, Zaiyi Chen +3

    cs.LGcs.CVstat.MLarXiv:1908.06112v12019
  23. Deep Learning for Medical Image Processing: Overview, Challenges and Future

    Muhammad Imran Razzak, Saeeda Naz, Ahmad Zaib

    cs.CVarXiv:1704.06825v12017
  24. Better Models, Faster Training: Sigmoid Attention for single-cell Foundation Models

    Vijay Sadashivaiah, Georgios Dasoulas, Judith Mueller +1

    cs.LGq-bio.QMarXiv:2604.27124v12026
  25. Gemma: Open Models Based on Gemini Research and Technology

    Gemma Team, Thomas Mesnard, Cassidy Hardin +105

    cs.CLcs.AIarXiv:2403.08295v42024
  26. MDETR -- Modulated Detection for End-to-End Multi-Modal Understanding

    Aishwarya Kamath, Mannat Singh, Yann LeCun +3

    cs.CVcs.CLcs.LGarXiv:2104.12763v22021
  27. FDA: Fourier Domain Adaptation for Semantic Segmentation

    Yanchao Yang, Stefano Soatto

    cs.CVarXiv:2004.05498v12020
  28. Scaling Laws for Reward Model Overoptimization

    Leo Gao, John Schulman, Jacob Hilton

    cs.LGstat.MLarXiv:2210.10760v12022
  29. How to Explain Individual Classification Decisions

    David Baehrens, Timon Schroeter, Stefan Harmeling +3

    stat.MLcs.LGarXiv:0912.1128v12009
  30. RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

    Yan Duan, John Schulman, Xi Chen +3

    cs.AIcs.LGcs.NEarXiv:1611.02779v22016
  31. Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models

    Xingwei Tan, Marco Valentino, Mahmud Elahi Akhter +3

    cs.CLcs.AIarXiv:2604.27251v22026
  32. State-of-the-art Speech Recognition With Sequence-to-Sequence Models

    Chung-Cheng Chiu, Tara N. Sainath, Yonghui Wu +11

    cs.CLcs.SDeess.ASarXiv:1712.01769v62017
  33. Voxel R-CNN: Towards High Performance Voxel-based 3D Object Detection

    Jiajun Deng, Shaoshuai Shi, Peiwei Li +3

    cs.CVarXiv:2012.15712v22020
  34. Latent Retrieval for Weakly Supervised Open Domain Question Answering

    Kenton Lee, Ming-Wei Chang, Kristina Toutanova

    cs.CLarXiv:1906.00300v32019
  35. Large Language Models are not Fair Evaluators

    Peiyi Wang, Lei Li, Liang Chen +7

    cs.CLcs.AIcs.IRarXiv:2305.17926v22023
  36. Large Language Models Explore by Latent Distilling

    Yuanhao Zeng, Ao Lu, Lufei Li +3

    cs.CLcs.AIcs.LGarXiv:2604.24927v22026
  37. Vital nodes identification in complex networks

    Linyuan Lü, Duanbing Chen, Xiao-Long Ren +3

    physics.soc-phcs.SIarXiv:1607.01134v12016
  38. Safety Drift After Fine-Tuning: Evidence from High-Stakes Domains

    Emaan Bilal Khan, Amy Winecoff, Miranda Bogen +1

    cs.CYcs.SEarXiv:2604.24902v12026
  39. Deep Learning: A Critical Appraisal

    Gary Marcus

    cs.AIcs.LGstat.MLarXiv:1801.00631v12018
  40. RWKV: Reinventing RNNs for the Transformer Era

    Bo Peng, Eric Alcaide, Quentin Anthony +31

    cs.CLcs.AIarXiv:2305.13048v22023
  41. X2SAM: Any Segmentation in Images and Videos

    Hao Wang, Limeng Qiao, Chi Zhang +4

    cs.CVcs.AIarXiv:2605.00891v12026
  42. Accelerating 3D Deep Learning with PyTorch3D

    Nikhila Ravi, Jeremy Reizenstein, David Novotny +4

    cs.CVcs.GRcs.LGarXiv:2007.08501v12020
  43. FcaNet: Frequency Channel Attention Networks

    Zequn Qin, Pengyi Zhang, Fei Wu +1

    cs.CVarXiv:2012.11879v42020
  44. Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models

    Jiayi Guo, Linqing Wang, Jiangshan Wang +6

    cs.CVarXiv:2604.25636v12026
  45. Importance Estimation for Neural Network Pruning

    Pavlo Molchanov, Arun Mallya, Stephen Tyree +2

    cs.LGcs.CVstat.MLarXiv:1906.10771v12019
  46. RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments

    Zaid Nasser, Mikhail Iumanov, Tianhao Li +3

    cs.CVarXiv:2604.26067v12026
  47. Prior-Aligned Data Cleaning for Tabular Foundation Models

    Laure Berti-Equille

    cs.LGcs.DBarXiv:2604.25154v12026
  48. Learning to Reconstruct 3D Human Pose and Shape via Model-fitting in the Loop

    Nikos Kolotouros, Georgios Pavlakos, Michael J. Black +1

    cs.CVarXiv:1909.12828v12019
  49. RULER: What's the Real Context Size of Your Long-Context Language Models?

    Cheng-Ping Hsieh, Simeng Sun, Samuel Kriman +5

    cs.CLarXiv:2404.06654v32024
  50. Pseudo-LiDAR from Visual Depth Estimation: Bridging the Gap in 3D Object Detection for Autonomous Driving

    Yan Wang, Wei-Lun Chao, Divyansh Garg +3

    cs.CVarXiv:1812.07179v62018
  51. LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale

    Tim Dettmers, Mike Lewis, Younes Belkada +1

    cs.LGcs.AIarXiv:2208.07339v22022
  52. FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption

    Yanting Wang, Chenlong Yin, Ying Chen +1

    cs.CRarXiv:2604.28157v12026
  53. Instruction-Guided Poetry Generation in Arabic and Its Dialects

    Abdelrahman Sadallah, Kareem Elozeiri, Mervat Abassy +5

    cs.CLcs.AIarXiv:2604.27766v12026
  54. Agentic Fusion of Large Atomic and Language Models to Accelerate Superconductor Discovery

    Mingze Li, Yu Rong, Songyou Li +16

    cs.LGcond-mat.mtrl-sciarXiv:2604.23758v32026
  55. Exploring the Limits of Language Modeling

    Rafal Jozefowicz, Oriol Vinyals, Mike Schuster +2

    cs.CLarXiv:1602.02410v22016
  56. FASH-iCNN: Making Editorial Fashion Identity Inspectable Through Multimodal CNN Probing

    Morayo Danielle Adeyemi, Ryan A. Rossi, Franck Dernoncourt

    cs.CVcs.HCcs.IRarXiv:2604.26186v12026
  57. Structural-RNN: Deep Learning on Spatio-Temporal Graphs

    Ashesh Jain, Amir R. Zamir, Silvio Savarese +1

    cs.CVcs.LGcs.NEarXiv:1511.05298v32015
  58. Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods

    Jieyu Zhao, Tianlu Wang, Mark Yatskar +2

    cs.CLcs.AIarXiv:1804.06876v12018
  59. Relief-Based Feature Selection: Introduction and Review

    Ryan J. Urbanowicz, Melissa Meeker, William LaCava +2

    cs.DScs.LGstat.MLarXiv:1711.08421v22017
  60. Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models

    Gongbo Zhang, Wen Wang, Ye Tian +1

    cs.CLcs.AIcs.LGarXiv:2604.26951v12026