Neural and Evolutionary Computing
Papers filed under cs.NE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.
Search paper metadata (including unsummarized papers)
1,141 to 1,200 of 1,364
PathNet: Evolution Channels Gradient Descent in Super Neural Networks
Chrisantha Fernando, Dylan Banarse, Charles Blundell +5
cs.NEcs.LGarXiv:1701.08734v12017Going Deeper in Facial Expression Recognition using Deep Neural Networks
Ali Mollahosseini, David Chan, Mohammad H. Mahoor
cs.NEcs.CVarXiv:1511.04110v12015Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures
Hengyuan Hu, Rui Peng, Yu-Wing Tai +1
cs.NEcs.CVcs.LGarXiv:1607.03250v12016Fast Algorithms for Convolutional Neural Networks
Andrew Lavin, Scott Gray
cs.NEcs.LGarXiv:1509.09308v22015Unity: A General Platform for Intelligent Agents
Arthur Juliani, Vincent-Pierre Berges, Ervin Teng +8
cs.LGcs.AIcs.NEarXiv:1809.02627v22018MADE: Masked Autoencoder for Distribution Estimation
Mathieu Germain, Karol Gregor, Iain Murray +1
cs.LGcs.NEstat.MLarXiv:1502.03509v22015"Zero-Shot" Super-Resolution using Deep Internal Learning
Assaf Shocher, Nadav Cohen, Michal Irani
cs.CVcs.LGcs.NEarXiv:1712.06087v12017Illuminating search spaces by mapping elites
Jean-Baptiste Mouret, Jeff Clune
cs.AIcs.NEcs.ROarXiv:1504.04909v12015Sparsity in Deep Learning: Pruning and growth for efficient inference and training in neural networks
Torsten Hoefler, Dan Alistarh, Tal Ben-Nun +2
cs.LGcs.AIcs.ARarXiv:2102.00554v12021Deep Predictive Coding Networks for Video Prediction and Unsupervised Learning
William Lotter, Gabriel Kreiman, David Cox
cs.LGcs.AIcs.CVarXiv:1605.08104v52016Visual7W: Grounded Question Answering in Images
Yuke Zhu, Oliver Groth, Michael Bernstein +1
cs.CVcs.LGcs.NEarXiv:1511.03416v42015GraphVAE: Towards Generation of Small Graphs Using Variational Autoencoders
Martin Simonovsky, Nikos Komodakis
cs.LGcs.CVcs.NEarXiv:1802.03480v12018Deep Big Simple Neural Nets Excel on Handwritten Digit Recognition
Dan Claudiu Ciresan, Ueli Meier, Luca Maria Gambardella +1
cs.NEcs.AIarXiv:1003.0358v12010PACEvolve: Enabling Long-Horizon Progress-Aware Consistent Evolution
Minghao Yan, Bo Peng, Benjamin Coleman +13
cs.NEcs.LGarXiv:2601.10657v22026Stochastic Pooling for Regularization of Deep Convolutional Neural Networks
Matthew D. Zeiler, Rob Fergus
cs.LGcs.NEstat.MLarXiv:1301.3557v12013Towards a Human-like Open-Domain Chatbot
Daniel Adiwardana, Minh-Thang Luong, David R. So +8
cs.CLcs.LGcs.NEarXiv:2001.09977v32020fastai: A Layered API for Deep Learning
Jeremy Howard, Sylvain Gugger
cs.LGcs.CVcs.NEarXiv:2002.04688v22020Cuckoo Search: Recent Advances and Applications
Xin-She Yang, Suash Deb
math.OCcs.NEnlin.AOarXiv:1408.5316v12014Med-BERT: pre-trained contextualized embeddings on large-scale structured electronic health records for disease prediction
Laila Rasmy, Yang Xiang, Ziqian Xie +2
cs.CLcs.LGcs.NEarXiv:2005.12833v12020Discriminative Unsupervised Feature Learning with Exemplar Convolutional Neural Networks
Alexey Dosovitskiy, Philipp Fischer, Jost Tobias Springenberg +2
cs.LGcs.CVcs.NEarXiv:1406.6909v22014Deep Complex Networks
Chiheb Trabelsi, Olexa Bilaniuk, Ying Zhang +7
cs.NEcs.LGarXiv:1705.09792v42017The NarrativeQA Reading Comprehension Challenge
Tomáš Kočiský, Jonathan Schwarz, Phil Blunsom +4
cs.CLcs.AIcs.NEarXiv:1712.07040v12017Training Deep Spiking Neural Networks using Backpropagation
Jun Haeng Lee, Tobi Delbruck, Michael Pfeiffer
cs.NEarXiv:1608.08782v12016How to Construct Deep Recurrent Neural Networks
Razvan Pascanu, Caglar Gulcehre, Kyunghyun Cho +1
cs.NEcs.LGstat.MLarXiv:1312.6026v52013Revisiting the Platonic Representation Hypothesis: An Aristotelian View
Fabian Gröger, Shuo Wen, Maria Brbić
cs.LGcs.AIcs.CVarXiv:2602.14486v22026Show Your Work: Scratchpads for Intermediate Computation with Language Models
Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari +9
cs.LGcs.NEarXiv:2112.00114v12021Universal statistical signatures of evolution in artificial intelligence architectures
Theodor Spiro
q-bio.PEcs.AIcs.CYarXiv:2604.10571v12026Event-triggered Implicit Perturbation for Zeroth-Order Fine-Tuning of Spiking Transformers
Tengteng Lei, Prabodh Katti, Rashi Dutt +5
cs.ARcs.LGcs.NEarXiv:2608.21223v12026What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search
Xinhao Zhang, Xi Chen, François Portet +1
cs.CLcs.NEarXiv:2604.19440v12026Uncertainty propagation in auto-regressive random neural network models
Janice Adams, Daniele Venturi
stat.MLcs.LGcs.NEarXiv:2608.20483v12026Training Deep Neural Networks on Noisy Labels with Bootstrapping
Scott Reed, Honglak Lee, Dragomir Anguelov +3
cs.CVcs.LGcs.NEarXiv:1412.6596v32014Resnet in Resnet: Generalizing Residual Architectures
Sasha Targ, Diogo Almeida, Kevin Lyman
cs.LGcs.CVcs.NEarXiv:1603.08029v12016Fine-Grain GPU Parallelization of the Generalized Partition Crossover for Large-Scale Traveling Salesman Problems
Swetha Varadarajan, Darrell Whitley
cs.AIcs.NEarXiv:2608.21233v12026Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition
Haşim Sak, Andrew Senior, Françoise Beaufays
cs.NEcs.CLcs.LGarXiv:1402.1128v12014Incremental Network Quantization: Towards Lossless CNNs with Low-Precision Weights
Aojun Zhou, Anbang Yao, Yiwen Guo +2
cs.CVcs.AIcs.NEarXiv:1702.03044v22017Bayesian SegNet: Model Uncertainty in Deep Convolutional Encoder-Decoder Architectures for Scene Understanding
Alex Kendall, Vijay Badrinarayanan, Roberto Cipolla
cs.CVcs.NEarXiv:1511.02680v22015A Hierarchical Latent Variable Encoder-Decoder Model for Generating Dialogues
Iulian Vlad Serban, Alessandro Sordoni, Ryan Lowe +4
cs.CLcs.AIcs.LGarXiv:1605.06069v32016TUDataset: A collection of benchmark datasets for learning with graphs
Christopher Morris, Nils M. Kriege, Franka Bause +3
cs.LGcs.NEstat.MLarXiv:2007.08663v12020Visualizing and Understanding Recurrent Networks
Andrej Karpathy, Justin Johnson, Li Fei-Fei
cs.LGcs.CLcs.NEarXiv:1506.02078v22015Dynamic Network Surgery for Efficient DNNs
Yiwen Guo, Anbang Yao, Yurong Chen
cs.NEcs.CVcs.LGarXiv:1608.04493v22016Neural Responding Machine for Short-Text Conversation
Lifeng Shang, Zhengdong Lu, Hang Li
cs.CLcs.AIcs.NEarXiv:1503.02364v22015Robots that can adapt like animals
Antoine Cully, Jeff Clune, Danesh Tarapore +1
cs.ROcs.AIcs.LGarXiv:1407.3501v42014A Network-based End-to-End Trainable Task-oriented Dialogue System
Tsung-Hsien Wen, David Vandyke, Nikola Mrksic +5
cs.CLcs.AIcs.NEarXiv:1604.04562v32016Regularizing and Optimizing LSTM Language Models
Stephen Merity, Nitish Shirish Keskar, Richard Socher
cs.CLcs.LGcs.NEarXiv:1708.02182v12017Deep Double Descent: Where Bigger Models and More Data Hurt
Preetum Nakkiran, Gal Kaplun, Yamini Bansal +3
cs.LGcs.CVcs.NEarXiv:1912.02292v12019Long Short-Term Memory-Networks for Machine Reading
Jianpeng Cheng, Li Dong, Mirella Lapata
cs.CLcs.NEarXiv:1601.06733v72016Activation Functions in Deep Learning: A Comprehensive Survey and Benchmark
Shiv Ram Dubey, Satish Kumar Singh, Bidyut Baran Chaudhuri
cs.LGcs.NEarXiv:2109.14545v32021Data Augmentation Generative Adversarial Networks
Antreas Antoniou, Amos Storkey, Harrison Edwards
stat.MLcs.CVcs.LGarXiv:1711.04340v32017RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning
Yan Duan, John Schulman, Xi Chen +3
cs.AIcs.LGcs.NEarXiv:1611.02779v22016Structural-RNN: Deep Learning on Spatio-Temporal Graphs
Ashesh Jain, Amir R. Zamir, Silvio Savarese +1
cs.CVcs.LGcs.NEarXiv:1511.05298v32015End-to-End Attention-based Large Vocabulary Speech Recognition
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk +2
cs.CLcs.AIcs.LGarXiv:1508.04395v22015Neural Module Networks
Jacob Andreas, Marcus Rohrbach, Trevor Darrell +1
cs.CVcs.CLcs.LGarXiv:1511.02799v42015A disciplined approach to neural network hyper-parameters: Part 1 -- learning rate, batch size, momentum, and weight decay
Leslie N. Smith
cs.LGcs.CVcs.NEarXiv:1803.09820v22018Generating Images with Perceptual Similarity Metrics based on Deep Networks
Alexey Dosovitskiy, Thomas Brox
cs.LGcs.CVcs.NEarXiv:1602.02644v22016Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results
Antti Tarvainen, Harri Valpola
cs.NEcs.LGstat.MLarXiv:1703.01780v62017Learning to Act and Cooperate for Distributed Black-Box Consensus Optimization
Zi-Bo Qin, Feng-Feng Wei, Tai-You Chen +1
cs.MAcs.NEarXiv:2605.00691v12026Ask Me Anything: Dynamic Memory Networks for Natural Language Processing
Ankit Kumar, Ozan Irsoy, Peter Ondruska +6
cs.CLcs.LGcs.NEarXiv:1506.07285v52015Compressing Deep Convolutional Networks using Vector Quantization
Yunchao Gong, Liu Liu, Ming Yang +1
cs.CVcs.LGcs.NEarXiv:1412.6115v12014Out-of-Distribution Generalization via Risk Extrapolation (REx)
David Krueger, Ethan Caballero, Joern-Henrik Jacobsen +5
cs.LGcs.AIcs.NEarXiv:2003.00688v52020Tensor field networks: Rotation- and translation-equivariant neural networks for 3D point clouds
Nathaniel Thomas, Tess Smidt, Steven Kearnes +4
cs.LGcs.AIcs.CVarXiv:1802.08219v32018