Every paper with a summary
Every arXiv paper Paperlayer has summarized so far. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
4,081 to 4,140 of 61,263
FreeSeg: Unified, Universal and Open-Vocabulary Image Segmentation
Jie Qin, Jie Wu, Pengxiang Yan +8
cs.CVarXiv:2303.17225v12023To See is to Believe: Prompting GPT-4V for Better Visual Instruction Tuning
Junke Wang, Lingchen Meng, Zejia Weng +3
cs.CVarXiv:2311.07574v22023CoPa: General Robotic Manipulation through Spatial Constraints of Parts with Foundation Models
Haoxu Huang, Fanqi Lin, Yingdong Hu +2
cs.ROarXiv:2403.08248v12024Phases in a class of associative memories via hidden neurons
Toshihiro Ota, Masato Taki
cs.LGcond-mat.dis-nncs.NEarXiv:2609.10976v12026The Big Data Bootstrap
Ariel Kleiner, Ameet Talwalkar, Purnamrita Sarkar +1
cs.LGstat.MLarXiv:1206.6415v12012P3: Toward Privacy-Preserving Photo Sharing
Moo-Ryong Ra, Ramesh Govindan, Antonio Ortega
cs.CRcs.MMarXiv:1302.5062v12013Parallel Tensor Compression for Large-Scale Scientific Data
Woody Austin, Grey Ballard, Tamara G. Kolda
math.NAcs.DCarXiv:1510.06689v22015Blockchain and Trusted Computing: Problems, Pitfalls, and a Solution for Hyperledger Fabric
Marcus Brandenburger, Christian Cachin, Rüdiger Kapitza +1
cs.DCcs.CRarXiv:1805.08541v12018"HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
Lingyao Li, Lizhou Fan, Shubham Atreja +1
cs.CLcs.AIcs.HCarXiv:2304.10619v12023SwinLSTM:Improving Spatiotemporal Prediction Accuracy using Swin Transformer and LSTM
Song Tang, Chuang Li, Pu Zhang +1
cs.CVcs.AIarXiv:2308.09891v22023Audio Deepfake Detection: A Survey
Jiangyan Yi, Chenglong Wang, Jianhua Tao +3
cs.SDeess.ASarXiv:2308.14970v12023SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension
Bohao Li, Yuying Ge, Yi Chen +3
cs.CVarXiv:2404.16790v12024Wasserstein Distributionally Robust Stochastic Control: A Data-Driven Approach
Insoon Yang
math.OCeess.SYarXiv:1812.09808v42018USB: A Unified Semi-supervised Learning Benchmark for Classification
Yidong Wang, Hao Chen, Yue Fan +19
cs.LGcs.AIcs.CVarXiv:2208.07204v22022Latent Constraints: Learning to Generate Conditionally from Unconditional Generative Models
Jesse Engel, Matthew Hoffman, Adam Roberts
cs.LGcs.NEstat.MLarXiv:1711.05772v22017Relatively Smart II: Tractable or Semi-Supervised Instance-Optimal Learning
Shaddin Dughmi, Alireza F. Pour
cs.LGstat.MLarXiv:2609.10886v12026SMT 2.0: A Surrogate Modeling Toolbox with a focus on Hierarchical and Mixed Variables Gaussian Processes
Paul Saves, Remi Lafage, Nathalie Bartoli +6
cs.LGcs.MSmath.OCarXiv:2305.13998v52023Demystifying Fixed k-Nearest Neighbor Information Estimators
Weihao Gao, Sewoong Oh, Pramod Viswanath
cs.LGcs.ITstat.MLarXiv:1604.03006v22016The Euclid mission design
Giuseppe D Racca, Rene Laureijs, Luca Stagnaro +39
astro-ph.IMarXiv:1610.05508v12016Certifying Lower Bounds for Risk-Sensitive Reinforcement Learning under Adversarial State Perturbations
Tong Li, Saunak Kumar Panda, Yisha Xiang
cs.LGmath.OCarXiv:2609.10866v12026Backpropagation through Signal Temporal Logic Specifications: Infusing Logical Structure into Gradient-Based Methods
Karen Leung, Nikos Aréchiga, Marco Pavone
eess.SYcs.CLcs.LOarXiv:2008.00097v32020Lattice Strategies for the Dirty Multiple Access Channel
Tal Philosof, Ram Zamir, Uri Erez +1
cs.ITarXiv:0904.1892v12009A Survey of Automatic Facial Micro-expression Analysis: Databases, Methods and Challenges
Yee-Hui Oh, John See, Anh Cat Le Ngo +2
cs.CVcs.MMarXiv:1806.05781v12018Imagination improves Multimodal Translation
Desmond Elliott, Ákos Kádár
cs.CLcs.CVarXiv:1705.04350v22017Cooperative Training of Descriptor and Generator Networks
Jianwen Xie, Yang Lu, Ruiqi Gao +2
stat.MLcs.CVarXiv:1609.09408v32016DR-LabStack: Design and Implementation of a Clinician-Facing Web System for Diabetic Retinopathy Prediction
Yingfan Xu, Tieming Liu, Ye Liang
cs.LGcs.SEarXiv:2609.10796v12026M6: A Chinese Multimodal Pretrainer
Junyang Lin, Rui Men, An Yang +22
cs.CLarXiv:2103.00823v42021RiVaT-Fuse: Reliability-Calibrated Variational Tensor Fusion for Multimodal Prediction under Modality Uncertainty
Yingfan Xu, Tieming Liu, Ye Liang +1
cs.LGcs.CVarXiv:2609.10798v12026Unbiased Teacher v2: Semi-supervised Object Detection for Anchor-free and Anchor-based Detectors
Yen-Cheng Liu, Chih-Yao Ma, Zsolt Kira
cs.CVcs.LGarXiv:2206.09500v12022Hybrid Relay-Reflecting Intelligent Surface-Assisted Wireless Communication
Nhan Thanh Nguyen, Quang-Doanh Vu, Kyungchun Lee +1
eess.SParXiv:2103.03900v12021Counterfactual Marginalisation: Framework for Evaluating Robustness to Nuisance Variables
Yasin Ibrahim, Hermione Warr, Robin J. Evans +1
cs.LGcs.AIarXiv:2609.10778v12026Communication-Efficient Accurate Statistical Estimation
Jianqing Fan, Yongyi Guo, Kaizheng Wang
stat.MLcs.LGmath.OCarXiv:1906.04870v22019Deep Multitask Learning for Semantic Dependency Parsing
Hao Peng, Sam Thomson, Noah A. Smith
cs.CLarXiv:1704.06855v22017MultiSports: A Multi-Person Video Dataset of Spatio-Temporally Localized Sports Actions
Yixuan Li, Lei Chen, Runyu He +3
cs.CVarXiv:2105.07404v22021Multi-Cue Zero-Shot Learning with Strong Supervision
Zeynep Akata, Mateusz Malinowski, Mario Fritz +1
cs.CVarXiv:1603.08754v12016A Bellman Optimality Equation for Plasticity
Jeremy Lucas, Doina Precup
cs.LGarXiv:2609.10776v12026Graph Neural Networks with Continual Learning for Fake News Detection from Social Media
Yi Han, Shanika Karunasekera, Christopher Leckie
cs.SIcs.LGarXiv:2007.03316v22020Language Embedded Radiance Fields for Zero-Shot Task-Oriented Grasping
Adam Rashid, Satvik Sharma, Chung Min Kim +4
cs.ROcs.CVarXiv:2309.07970v22023PhyScene: Physically Interactable 3D Scene Synthesis for Embodied AI
Yandan Yang, Baoxiong Jia, Peiyuan Zhi +1
cs.CVcs.AIcs.LGarXiv:2404.09465v22024Learning Orthogonal Multi-Index Models Beyond Small Initialization: Incremental Learning, Competitive Dynamics and Symmetry
Mo Zhou, Weihang Xu, Simon S. Du +1
cs.LGstat.MLarXiv:2609.10879v12026Isotropic PCA and Affine-Invariant Clustering
S. Charles Brubaker, Santosh S. Vempala
cs.LGcs.CGarXiv:0804.3575v22008A Market-Inspired Approach for Intersection Management in Urban Road Traffic Networks
Matteo Vasirani, Sascha Ossowski
cs.GTcs.MAarXiv:1401.5851v12014Contextual Parameter Generation for Universal Neural Machine Translation
Emmanouil Antonios Platanios, Mrinmaya Sachan, Graham Neubig +1
cs.CLcs.LGstat.MLarXiv:1808.08493v12018Optimal Distributed Subsampling for Maximum Quasi-Likelihood Estimators with Massive Data
Jun Yu, HaiYing Wang, Mingyao Ai +1
stat.MEcs.DCstat.COarXiv:2005.10435v32020Time-Sensitive Networking in IEEE 802.11be: On the Way to Low-latency WiFi 7
Toni Adame, Marc Carrascosa, Boris Bellalta
cs.NIarXiv:1912.06086v22019Deep Generative Learning via Schrödinger Bridge
Gefei Wang, Yuling Jiao, Qian Xu +2
cs.LGcs.CVarXiv:2106.10410v22021Language Models that Seek for Knowledge: Modular Search & Generation for Dialogue and Prompt Completion
Kurt Shuster, Mojtaba Komeili, Leonard Adolphs +3
cs.CLcs.AIarXiv:2203.13224v22022FedPrompt: Communication-Efficient and Privacy Preserving Prompt Tuning in Federated Learning
Haodong Zhao, Wei Du, Fangqi Li +2
cs.LGcs.AIcs.CRarXiv:2208.12268v32022Flow Duality and Source Geometry for Categorical Generation
Etrit Haxholli
cs.LGstat.MLarXiv:2609.10863v12026Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Fu-Yun Wang, Wenshuo Chen, Guanglu Song +3
cs.CVarXiv:2305.18264v12023ChatTime: A Unified Multimodal Time Series Foundation Model Bridging Numerical and Textual Data
Chengsen Wang, Qi Qi, Jingyu Wang +5
cs.CLcs.LGarXiv:2412.11376v12024Answering Table Queries on the Web using Column Keywords
Rakesh Pimplikar, Sunita Sarawagi
cs.DBarXiv:1207.0132v12012Processing and classifying bird songs using wavelet techniques and supervised learning
Laura Lucia Dominguez Barrios, Fidel Aniano Causil Barrios, Alex Rodrigo dos Santos Sousa +1
cs.LGstat.MEarXiv:2609.10826v12026GenArtist: Multimodal LLM as an Agent for Unified Image Generation and Editing
Zhenyu Wang, Aoxue Li, Zhenguo Li +1
cs.CVarXiv:2407.05600v22024SparCML: High-Performance Sparse Communication for Machine Learning
Cedric Renggli, Saleh Ashkboos, Mehdi Aghagolzadeh +2
cs.DCstat.MLarXiv:1802.08021v32018Policies Modulating Trajectory Generators
Atil Iscen, Ken Caluwaerts, Jie Tan +4
cs.ROcs.AIcs.LGarXiv:1910.02812v12019A joint model of unpaired data from scRNA-seq and spatial transcriptomics for imputing missing gene expression measurements
Romain Lopez, Achille Nazaret, Maxime Langevin +4
cs.LGq-bio.GNstat.MLarXiv:1905.02269v12019The Consensus Number of a Cryptocurrency (Extended Version)
Rachid Guerraoui, Petr Kuznetsov, Matteo Monti +2
cs.DCarXiv:1906.05574v12019Text-to-Viz: Automatic Generation of Infographics from Proportion-Related Natural Language Statements
Weiwei Cui, Xiaoyu Zhang, Yun Wang +6
cs.HCarXiv:1907.09091v12019Pano: Optimizing 360° Video Streaming with a Better Understanding of Quality Perception
Yu Guan, Chengyuan Zheng, Zongming Guo +2
cs.MMarXiv:1911.04139v12019