Neural and Evolutionary Computing

Papers filed under cs.NE on arXiv, each one already summarized by Paperlayer. Open any of them to read the summary beside the original PDF, with every point linked to the line, figure, or table it came from.

Search paper metadata (including unsummarized papers)

1,141 to 1,200 of 1,364

  1. PathNet: Evolution Channels Gradient Descent in Super Neural Networks

    Chrisantha Fernando, Dylan Banarse, Charles Blundell +5

    cs.NEcs.LGarXiv:1701.08734v12017
  2. Going Deeper in Facial Expression Recognition using Deep Neural Networks

    Ali Mollahosseini, David Chan, Mohammad H. Mahoor

    cs.NEcs.CVarXiv:1511.04110v12015
  3. Network Trimming: A Data-Driven Neuron Pruning Approach towards Efficient Deep Architectures

    Hengyuan Hu, Rui Peng, Yu-Wing Tai +1

    cs.NEcs.CVcs.LGarXiv:1607.03250v12016
  4. Fast Algorithms for Convolutional Neural Networks

    Andrew Lavin, Scott Gray

    cs.NEcs.LGarXiv:1509.09308v22015
  5. Unity: A General Platform for Intelligent Agents

    Arthur Juliani, Vincent-Pierre Berges, Ervin Teng +8

    cs.LGcs.AIcs.NEarXiv:1809.02627v22018
  6. MADE: Masked Autoencoder for Distribution Estimation

    Mathieu Germain, Karol Gregor, Iain Murray +1

    cs.LGcs.NEstat.MLarXiv:1502.03509v22015
  7. "Zero-Shot" Super-Resolution using Deep Internal Learning

    Assaf Shocher, Nadav Cohen, Michal Irani

    cs.CVcs.LGcs.NEarXiv:1712.06087v12017
  8. Illuminating search spaces by mapping elites

    Jean-Baptiste Mouret, Jeff Clune

    cs.AIcs.NEcs.ROarXiv:1504.04909v12015
  9. Sparsity in Deep Learning: Pruning and growth for efficient inference and training in neural networks

    Torsten Hoefler, Dan Alistarh, Tal Ben-Nun +2

    cs.LGcs.AIcs.ARarXiv:2102.00554v12021
  10. Deep Predictive Coding Networks for Video Prediction and Unsupervised Learning

    William Lotter, Gabriel Kreiman, David Cox

    cs.LGcs.AIcs.CVarXiv:1605.08104v52016
  11. Visual7W: Grounded Question Answering in Images

    Yuke Zhu, Oliver Groth, Michael Bernstein +1

    cs.CVcs.LGcs.NEarXiv:1511.03416v42015
  12. GraphVAE: Towards Generation of Small Graphs Using Variational Autoencoders

    Martin Simonovsky, Nikos Komodakis

    cs.LGcs.CVcs.NEarXiv:1802.03480v12018
  13. Deep Big Simple Neural Nets Excel on Handwritten Digit Recognition

    Dan Claudiu Ciresan, Ueli Meier, Luca Maria Gambardella +1

    cs.NEcs.AIarXiv:1003.0358v12010
  14. PACEvolve: Enabling Long-Horizon Progress-Aware Consistent Evolution

    Minghao Yan, Bo Peng, Benjamin Coleman +13

    cs.NEcs.LGarXiv:2601.10657v22026
  15. Stochastic Pooling for Regularization of Deep Convolutional Neural Networks

    Matthew D. Zeiler, Rob Fergus

    cs.LGcs.NEstat.MLarXiv:1301.3557v12013
  16. Towards a Human-like Open-Domain Chatbot

    Daniel Adiwardana, Minh-Thang Luong, David R. So +8

    cs.CLcs.LGcs.NEarXiv:2001.09977v32020
  17. fastai: A Layered API for Deep Learning

    Jeremy Howard, Sylvain Gugger

    cs.LGcs.CVcs.NEarXiv:2002.04688v22020
  18. Cuckoo Search: Recent Advances and Applications

    Xin-She Yang, Suash Deb

    math.OCcs.NEnlin.AOarXiv:1408.5316v12014
  19. Med-BERT: pre-trained contextualized embeddings on large-scale structured electronic health records for disease prediction

    Laila Rasmy, Yang Xiang, Ziqian Xie +2

    cs.CLcs.LGcs.NEarXiv:2005.12833v12020
  20. Discriminative Unsupervised Feature Learning with Exemplar Convolutional Neural Networks

    Alexey Dosovitskiy, Philipp Fischer, Jost Tobias Springenberg +2

    cs.LGcs.CVcs.NEarXiv:1406.6909v22014
  21. Deep Complex Networks

    Chiheb Trabelsi, Olexa Bilaniuk, Ying Zhang +7

    cs.NEcs.LGarXiv:1705.09792v42017
  22. The NarrativeQA Reading Comprehension Challenge

    Tomáš Kočiský, Jonathan Schwarz, Phil Blunsom +4

    cs.CLcs.AIcs.NEarXiv:1712.07040v12017
  23. Training Deep Spiking Neural Networks using Backpropagation

    Jun Haeng Lee, Tobi Delbruck, Michael Pfeiffer

    cs.NEarXiv:1608.08782v12016
  24. How to Construct Deep Recurrent Neural Networks

    Razvan Pascanu, Caglar Gulcehre, Kyunghyun Cho +1

    cs.NEcs.LGstat.MLarXiv:1312.6026v52013
  25. Revisiting the Platonic Representation Hypothesis: An Aristotelian View

    Fabian Gröger, Shuo Wen, Maria Brbić

    cs.LGcs.AIcs.CVarXiv:2602.14486v22026
  26. Show Your Work: Scratchpads for Intermediate Computation with Language Models

    Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari +9

    cs.LGcs.NEarXiv:2112.00114v12021
  27. Universal statistical signatures of evolution in artificial intelligence architectures

    Theodor Spiro

    q-bio.PEcs.AIcs.CYarXiv:2604.10571v12026
  28. Event-triggered Implicit Perturbation for Zeroth-Order Fine-Tuning of Spiking Transformers

    Tengteng Lei, Prabodh Katti, Rashi Dutt +5

    cs.ARcs.LGcs.NEarXiv:2608.21223v12026
  29. What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search

    Xinhao Zhang, Xi Chen, François Portet +1

    cs.CLcs.NEarXiv:2604.19440v12026
  30. Uncertainty propagation in auto-regressive random neural network models

    Janice Adams, Daniele Venturi

    stat.MLcs.LGcs.NEarXiv:2608.20483v12026
  31. Training Deep Neural Networks on Noisy Labels with Bootstrapping

    Scott Reed, Honglak Lee, Dragomir Anguelov +3

    cs.CVcs.LGcs.NEarXiv:1412.6596v32014
  32. Resnet in Resnet: Generalizing Residual Architectures

    Sasha Targ, Diogo Almeida, Kevin Lyman

    cs.LGcs.CVcs.NEarXiv:1603.08029v12016
  33. Fine-Grain GPU Parallelization of the Generalized Partition Crossover for Large-Scale Traveling Salesman Problems

    Swetha Varadarajan, Darrell Whitley

    cs.AIcs.NEarXiv:2608.21233v12026
  34. Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition

    Haşim Sak, Andrew Senior, Françoise Beaufays

    cs.NEcs.CLcs.LGarXiv:1402.1128v12014
  35. Incremental Network Quantization: Towards Lossless CNNs with Low-Precision Weights

    Aojun Zhou, Anbang Yao, Yiwen Guo +2

    cs.CVcs.AIcs.NEarXiv:1702.03044v22017
  36. Bayesian SegNet: Model Uncertainty in Deep Convolutional Encoder-Decoder Architectures for Scene Understanding

    Alex Kendall, Vijay Badrinarayanan, Roberto Cipolla

    cs.CVcs.NEarXiv:1511.02680v22015
  37. A Hierarchical Latent Variable Encoder-Decoder Model for Generating Dialogues

    Iulian Vlad Serban, Alessandro Sordoni, Ryan Lowe +4

    cs.CLcs.AIcs.LGarXiv:1605.06069v32016
  38. TUDataset: A collection of benchmark datasets for learning with graphs

    Christopher Morris, Nils M. Kriege, Franka Bause +3

    cs.LGcs.NEstat.MLarXiv:2007.08663v12020
  39. Visualizing and Understanding Recurrent Networks

    Andrej Karpathy, Justin Johnson, Li Fei-Fei

    cs.LGcs.CLcs.NEarXiv:1506.02078v22015
  40. Dynamic Network Surgery for Efficient DNNs

    Yiwen Guo, Anbang Yao, Yurong Chen

    cs.NEcs.CVcs.LGarXiv:1608.04493v22016
  41. Neural Responding Machine for Short-Text Conversation

    Lifeng Shang, Zhengdong Lu, Hang Li

    cs.CLcs.AIcs.NEarXiv:1503.02364v22015
  42. Robots that can adapt like animals

    Antoine Cully, Jeff Clune, Danesh Tarapore +1

    cs.ROcs.AIcs.LGarXiv:1407.3501v42014
  43. A Network-based End-to-End Trainable Task-oriented Dialogue System

    Tsung-Hsien Wen, David Vandyke, Nikola Mrksic +5

    cs.CLcs.AIcs.NEarXiv:1604.04562v32016
  44. Regularizing and Optimizing LSTM Language Models

    Stephen Merity, Nitish Shirish Keskar, Richard Socher

    cs.CLcs.LGcs.NEarXiv:1708.02182v12017
  45. Deep Double Descent: Where Bigger Models and More Data Hurt

    Preetum Nakkiran, Gal Kaplun, Yamini Bansal +3

    cs.LGcs.CVcs.NEarXiv:1912.02292v12019
  46. Long Short-Term Memory-Networks for Machine Reading

    Jianpeng Cheng, Li Dong, Mirella Lapata

    cs.CLcs.NEarXiv:1601.06733v72016
  47. Activation Functions in Deep Learning: A Comprehensive Survey and Benchmark

    Shiv Ram Dubey, Satish Kumar Singh, Bidyut Baran Chaudhuri

    cs.LGcs.NEarXiv:2109.14545v32021
  48. Data Augmentation Generative Adversarial Networks

    Antreas Antoniou, Amos Storkey, Harrison Edwards

    stat.MLcs.CVcs.LGarXiv:1711.04340v32017
  49. RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

    Yan Duan, John Schulman, Xi Chen +3

    cs.AIcs.LGcs.NEarXiv:1611.02779v22016
  50. Structural-RNN: Deep Learning on Spatio-Temporal Graphs

    Ashesh Jain, Amir R. Zamir, Silvio Savarese +1

    cs.CVcs.LGcs.NEarXiv:1511.05298v32015
  51. End-to-End Attention-based Large Vocabulary Speech Recognition

    Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk +2

    cs.CLcs.AIcs.LGarXiv:1508.04395v22015
  52. Neural Module Networks

    Jacob Andreas, Marcus Rohrbach, Trevor Darrell +1

    cs.CVcs.CLcs.LGarXiv:1511.02799v42015
  53. A disciplined approach to neural network hyper-parameters: Part 1 -- learning rate, batch size, momentum, and weight decay

    Leslie N. Smith

    cs.LGcs.CVcs.NEarXiv:1803.09820v22018
  54. Generating Images with Perceptual Similarity Metrics based on Deep Networks

    Alexey Dosovitskiy, Thomas Brox

    cs.LGcs.CVcs.NEarXiv:1602.02644v22016
  55. Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results

    Antti Tarvainen, Harri Valpola

    cs.NEcs.LGstat.MLarXiv:1703.01780v62017
  56. Learning to Act and Cooperate for Distributed Black-Box Consensus Optimization

    Zi-Bo Qin, Feng-Feng Wei, Tai-You Chen +1

    cs.MAcs.NEarXiv:2605.00691v12026
  57. Ask Me Anything: Dynamic Memory Networks for Natural Language Processing

    Ankit Kumar, Ozan Irsoy, Peter Ondruska +6

    cs.CLcs.LGcs.NEarXiv:1506.07285v52015
  58. Compressing Deep Convolutional Networks using Vector Quantization

    Yunchao Gong, Liu Liu, Ming Yang +1

    cs.CVcs.LGcs.NEarXiv:1412.6115v12014
  59. Out-of-Distribution Generalization via Risk Extrapolation (REx)

    David Krueger, Ethan Caballero, Joern-Henrik Jacobsen +5

    cs.LGcs.AIcs.NEarXiv:2003.00688v52020
  60. Tensor field networks: Rotation- and translation-equivariant neural networks for 3D point clouds

    Nathaniel Thomas, Tess Smidt, Steven Kearnes +4

    cs.LGcs.AIcs.CVarXiv:1802.08219v32018