Skip to content
— CH. 1 · INTRODUCTION —

Neural network (machine learning)

12 min listen · Ch. 1 of 8
8 sections
  • Neural networks have become the engine behind voice assistants, cancer diagnoses, deepfake videos, and the large language models now reshaping how people interact with machines. But the idea behind them is older, stranger, and more biologically grounded than most people realize. At its core, a neural network is a computational model inspired by the structure of the brain itself. Neurons, synapses, signals, weights: the vocabulary of neuroscience was consciously transplanted into mathematics and software. How did that transplant take hold? What kept it from working for decades? And what finally unlocked its extraordinary power?

  • Warren McCulloch and Walter Pitts, in 1943, published the first mathematical model of an artificial neuron capable of representing logical functions. Their paper considered neural networks that contain cycles and noted that the current activity of such networks could be affected by activity indefinitely far in the past. That single observation planted the seed for what would later become recurrent networks.

    The connection to biological structure was not merely metaphorical. Each artificial neuron receives numerical inputs, weights them, sums them with a bias term, and then passes the result through a nonlinear activation function to produce an output. The edges connecting neurons model the synapses in the brain. The layered structure, with an input layer, hidden layers, and an output layer, mirrors the way biological neural circuits process information from the sensory periphery inward.

    D. O. Hebb, in the late 1940s, proposed a learning hypothesis based on neural plasticity. That principle, known as Hebbian learning, held that connections between neurons strengthen when those neurons fire together. It became the basis for early experiments including Rosenblatt's perceptron and the Hopfield network. The brain's architecture was not just an inspiration; it was an active blueprint that researchers kept returning to, even as the field grew increasingly mathematical.

  • Frank Rosenblatt, a psychologist, described the perceptron in 1958. It was one of the first implemented neural networks and was funded by the United States Office of Naval Research. The announcement triggered public excitement, and the US government drastically increased funding for the field. Researchers spoke optimistically about perceptrons eventually emulating human intelligence, a period later called "the Golden Age of AI".

    That confidence collapsed. Marvin Minsky and Seymour Papert published a book called Perceptrons that highlighted a fundamental flaw: single-layer perceptrons could only solve linearly separable problems. They specifically emphasized that perceptrons were incapable of processing the exclusive-or circuit. The book deflated interest through the late 1960s and into the 1970s.

    What the critics missed is that work on deeper networks had already been underway. Alexey Ivakhnenko and Valentin Lapa published the first working deep learning algorithm in the Soviet Union in 1965, a method they regarded as a form of polynomial regression generalizing Rosenblatt's perceptron. A 1971 paper described a network with eight layers trained by this method. Shun'ichi Amari published the first deep learning multilayer perceptron trained by stochastic gradient descent in 1967. The Minsky-Papert critique was simply irrelevant to these deeper architectures, but the damage to funding and public confidence had already been done.

  • Interest in neural networks revived during the 1980s because of a new algorithm: backpropagation. The algorithm works by propagating error gradients backward from the output layer through to the input layer, adjusting the weights at each step. This made it possible to train multi-layer networks efficiently for the first time.

    The history of the algorithm is tangled. Henry J. Kelley developed a precursor in 1960 within control theory. Seppo Linnainmaa published the modern form in his master's thesis in 1970. G. M. Ostrovski et al. republished it in 1971. Paul Werbos applied it explicitly to neural networks in 1982. David E. Rumelhart and colleagues popularized it in 1986 but did not cite the original work. The terminology "back-propagating errors" had actually been introduced by Rosenblatt himself as early as 1962, though he did not describe how to implement it. Backpropagation is mathematically an efficient application of the chain rule, which Gottfried Wilhelm Leibniz derived in 1673.

    Once widely adopted, backpropagation opened the door to training the multi-layer architectures that single-layer critics had dismissed. It also set the stage for the GPU-accelerated training pipelines that would follow. From 1991 to 2015, computing power delivered by GPUs increased around a million-fold, enabling backpropagation to scale to networks of previously unimaginable depth.

  • Kunihiko Fukushima introduced the neocognitron in 1979, a deep convolutional architecture with convolutional layers, downsampling layers, weight replication, and max pooling. That structure became the foundation for convolutional neural networks, which proved transformative for computer vision.

    Yann LeCun and colleagues built LeNet in 1989 to recognize handwritten ZIP codes on mail; training required three days. By 1998, LeNet-5, a seven-level convolutional network, was being applied by banks to recognize numbers on checks digitized in 32 by 32 pixel images. In October 2012, AlexNet, built by Alex Krizhevsky, Ilya Sutskever, and Geoffrey Hinton, won the large-scale ImageNet competition by a significant margin over shallow machine learning methods. In 2011, DanNet had already achieved superhuman performance in a visual pattern recognition contest, outperforming traditional methods by a factor of three.

    Long short-term memory, or LSTM, addressed a different problem. Sepp Hochreiter's diploma thesis in 1991 identified the vanishing gradient problem that crippled deep recurrent networks. He and Jürgen Schmidhuber introduced LSTM, which set accuracy records across multiple domains. The forget gate, added in 1999, completed the architecture and made it the default choice for recurrent network design for years.

    Then in 2017, the paper "Attention Is All You Need" introduced the transformer architecture, which uses attention mechanisms to model long-range dependencies in data. Modern large language models including GPT, Gemini, Grok, DeepSeek, and Qwen are all built on this foundation.

  • Training a neural network means adjusting the weights of its connections until the network's outputs match desired targets closely enough to be useful. The process relies on a loss function that measures the degree of error between what the network predicts and what it should predict. As long as the loss function continues to decline, the network is improving. Training ends when additional observations no longer usefully reduce the cost.

    A hyperparameter is any configurable part of the network whose value is set before training begins: learning rate, batch size, number of layers, and number of nodes are all examples. The learning rate controls how large each corrective step is. A high learning rate shortens training time but sacrifices accuracy; a lower rate takes longer but allows greater precision. Adaptive learning rates, which increase or decrease as appropriate, help avoid oscillating between high and low weight values.

    Networks face a structural hazard called overfitting: when the network is so large relative to the training data that it memorizes the training examples rather than learning to generalize. Cross-validation and regularization techniques address this. In probabilistic frameworks, selecting a larger prior probability over simpler models reduces the tendency to overfit. Neural networks also require far more sample inputs than biological brains need to reach a given level of function. As of 2026, training a commercial large language model typically required hundreds of thousands of computers and cost tens of millions of dollars.

  • From 1988 onward, neural networks transformed the field of protein structure prediction, particularly when cascading networks began training on profiles produced by multiple sequence alignments. That application arrived well before the modern deep learning wave.

    In medicine, neural networks have been used to diagnose cancers and to distinguish highly invasive cancer cell lines from less invasive ones using only cell shape data. They have accelerated reliability analysis of infrastructure subject to natural disasters and predict settling in building foundations. In cybersecurity, they classify Android malware, identify domains belonging to threat actors, and detect URLs posing a security risk.

    Generative adversarial networks, introduced by Ian Goodfellow and colleagues in 2014, drove state-of-the-art image generation through 2018. Nvidia's StyleGAN, released in 2018 and based on the Progressive GAN by Tero Karras and colleagues, achieved excellent image quality and provoked widespread discussion about deepfakes. Diffusion models, emerging in 2015, eventually surpassed GANs in generative modeling, with systems such as DALL-E 2 and Stable Diffusion both arriving in 2022.

    DALL-E was trained on 650 million pairs of images and texts and can create artworks from user text descriptions. Companies including AIVA and Jukedeck have applied transformer architectures to generate original music. Film production companies have used neural networks to analyze the likely financial success of a film. In 2018, Amazon scrapped a recruiting tool after discovering the model favored men over women for software engineering roles because the training data itself reflected a workforce that was predominantly male; the system penalized resumes containing the word "woman" or the name of a women's college.

  • The multilayer perceptron is a universal function approximator, as established by the universal approximation theorem. A recurrent architecture with rational-valued weights can match the power of a universal Turing machine using a finite number of neurons and linear connections. Using irrational values for weights produces a machine with what researchers describe as super-Turing power.

    Despite this theoretical reach, neural networks face real constraints. The brain operates on roughly 20 watts of power. Training commercial transformer models in 2026 required data centers drawing hundreds of megawatts. The statistical properties of real-world data shift over time, a phenomenon called concept drift, which can cause a network's accuracy in deployment to diverge substantially from what was measured during training. Neural networks are also "black box" systems: their internal decision-making processes remain difficult to interpret, and they are vulnerable to adversarial examples that can cause incorrect predictions through deliberately crafted inputs.

    These concerns have spurred research into explainable artificial intelligence and hybrid models that combine neural learning with symbolic reasoning. Neuromorphic engineering has pursued a different path entirely, building non-von-Neumann chips designed to implement neural networks directly in circuitry; Alphabet introduced custom chips called Tensor Processing Units as one example of this direction. The vanishing gradient problem that Hochreiter identified in 1991 took years to solve; the interpretability problem has not yet found its equivalent breakthrough.

Common questions

What is a neural network in machine learning?

A neural network is a computational model inspired by the structure and functions of biological neural networks. It consists of connected artificial neurons arranged in layers, where signals travel from an input layer through hidden layers to an output layer, with weights on each connection adjusted during training to improve accuracy.

Who invented the first neural network?

Warren McCulloch and Walter Pitts published the first mathematical model of artificial neurons in 1943, capable of representing logical functions. Frank Rosenblatt described the perceptron in 1958, one of the first implemented neural networks, funded by the United States Office of Naval Research.

What is backpropagation and who developed it?

Backpropagation is an algorithm that trains neural networks by propagating error gradients backward from the output layer to the input layer to adjust connection weights. Seppo Linnainmaa published the modern form in his master's thesis in 1970; David E. Rumelhart and colleagues popularized it in 1986.

What is the difference between a convolutional neural network and a recurrent neural network?

Convolutional neural networks, whose deep architecture originated with Kunihiko Fukushima's neocognitron in 1979, are specialized for spatially structured data like images and are the essential tool for computer vision. Recurrent neural networks allow connections between neurons in the same or previous layers, making them suited for sequential data such as speech and time series; the LSTM architecture, introduced by Sepp Hochreiter and Jürgen Schmidhuber, became the default recurrent design after the forget gate was added in 1999.

What is the transformer architecture and why does it matter?

The transformer architecture was introduced in 2017 in the paper "Attention Is All You Need" and uses attention mechanisms to model long-range dependencies in data. Large language models including GPT, Gemini, Grok, DeepSeek, and Qwen are all built on this architecture.

How much does it cost to train a large language model in 2026?

As of 2026, training a commercial large language model such as GPT, Grok, or Gemini typically required hundreds of thousands of computers and cost tens of millions of dollars. Training these models requires data centers drawing hundreds of megawatts of power, compared to the roughly 20 watts the human brain consumes.

All sources

234 references cited across the entry

  1. 1Explained: Neural networksLarry Hardesty — MIT News Office — 14 April 2017
  2. 2BookComprehensive Biomedical PhysicsZ.R. Yang et al. — Elsevier — 2014
  3. 3BookPattern Recognition and Machine LearningChristopher M. Bishop — Springer — 17 August 2006
  4. 4JournalAttention Is All You NeedAshish Vaswani et al. — 2017
  5. 5BookNeural Networks for BabiesFerrie, C. — Sourcebooks — 2019
  6. 6JournalGauss and the Invention of Least SquaresStephen M. Stigler — 1981
  7. 7Annotated History of Modern AI and Deep LearningJuergen Schmidhuber — 2025-12-29
  8. 9JournalA logical calculus of the ideas immanent in nervous activityWarren S. McCulloch et al. — December 1943
  9. 10NewsRepresentation of Events in Nerve Nets and Finite AutomataS.C. Kleene — Princeton University Press — 1956
  10. 11BookThe Organization of BehaviorDonald Hebb — Taylor & Francis — 2005
  11. 12JournalSimulation of Self-Organizing Systems by Digital ComputerB.G. Farley — 1954
  12. 13JournalTests on a cell assembly theory of the action of the brain, using a large digital computerN. Rochester — 1956
  13. 14JournalThe Perceptron: A Probabilistic Model For Information Storage And Organization in the BrainF. Rosenblatt — 1958
  14. 15BookBeyond Regression: New Tools for Prediction and Analysis in the Behavioral SciencesP.J. Werbos — 1975
  15. 16JournalThe Perceptron—a perceiving and recognizing automatonFrank Rosenblatt — Cornell Aeronautical Laboratory — 1957
  16. 17JournalA Sociological Study of the Official History of the Perceptrons ControversyMikel Olazaran — 1996
  17. 18ThesisContributions to perceptron theoryR. D. Joseph — Cornell University — 1961
  18. 20BookArtificial Intelligence A Modern ApproachRussel — Pearson Education — 2010
  19. 22BookPerceptrons: An Introduction to Computational GeometryMarvin Minsky et al. — MIT Press — 1969
  20. 23BookCybernetics and Forecasting TechniquesAlexey G. Ivakhnenko et al. — American Elsevier Publishing Co. — 1967
  21. 25JournalPolynomial theory of complex systemsAlexey Ivakhnenko — 1971
  22. 26JournalA Stochastic Approximation MethodH. Robbins et al. — 1951
  23. 27JournalA theory of adaptive pattern classifiersShun'ichi Amari — 1967
  24. 28JournalVisual feature extraction by a multilayered network of analog threshold elementsK. Fukushima — 1969
  25. 29JournalNeural network with unbounded activation functions is universal approximatorSho Sonoda et al. — 2017
  26. 30Searching for Activation FunctionsPrajit Ramachandran et al. — 16 October 2017
  27. 31BookPerceptrons: An Introduction to Computational GeometryMarvin Minsky et al. — MIT Press — 1969
  28. 33JournalLearning representations by back-propagating errorsDavid Rumelhart et al. — 1986
  29. 35JournalGradient theory of optimal flight pathsHenry J. Kelley — 1960
  30. 36ThesisThe representation of the cumulative rounding error of an algorithm as a Taylor expansion of the local rounding errorsSeppo Linnainmaa — University of Helsinki — 1970
  31. 37JournalTaylor expansion of the accumulated rounding errorSeppo Linnainmaa — 1976
  32. 38Who Invented Backpropagation?Juergen Schmidhuber — IDSIA, Switzerland — 25 October 2014
  33. 39Applications of advances in nonlinear sensitivity analysisPaul J. Werbos — Springer-Verlag — 1982
  34. 41BookThe Roots of Backpropagation: From Ordered Derivatives to Neural Networks and Political ForecastingPaul J. Werbos — John Wiley & Sons — 1994
  35. 42JournalLearning representations by back-propagating errorsDavid E. Rumelhart et al. — October 1986
  36. 45JournalDeep Learning in Neural Networks: An OverviewJ. Schmidhuber — 2015
  37. 55JournalNeural networks and physical systems with emergent collective computational abilitiesJ. J. Hopfield — 1982
  38. 56JournalThe Importance of Cajal's and Lorente de Nó's Neuroscience to the Birth of CyberneticsJuan Manuel Espinosa-Sanchez et al. — 5 July 2023
  39. 60JournalThoughts on the relations between emotion and cognition.Richard S. Lazarus — November 1982
  40. 62JournalNeural Sequence ChunkersJürgen Schmidhuber — April 1991
  41. 66BookA Field Guide to Dynamical Recurrent NetworksS. Hochreiter — John Wiley & Sons — 15 January 2001
  42. 67JournalLong Short-Term MemorySepp Hochreiter et al. — 1 November 1997
  43. 68Book9th International Conference on Artificial Neural Networks: ICANN '99Felix Gers et al. — 1999
  44. 69JournalA learning algorithm for boltzmann machinesDavid H. Ackley et al. — 1 January 1985
  45. 70BookParallel Distributed Processing: Explorations in the Microstructure of Cognition, Volume 1: FoundationsPaul Smolensky — MIT Press — 1986
  46. 71JournalThe Helmholtz machine.Dayan Peter et al. — 1995
  47. 72JournalThe wake-sleep algorithm for unsupervised neural networksGeoffrey E. Hinton et al. — 26 May 1995
  48. 75JournalDeep, Big, Simple Neural Nets for Handwritten Digit RecognitionDan Claudiu Cireşan et al. — 21 September 2010
  49. 77BookAdvances in Neural Information Processing Systems 25Dan Ciresan et al. — Curran Associates, Inc. — 2012
  50. 78BookMedical Image Computing and Computer-Assisted Intervention – MICCAI 2013D. Ciresan et al. — 2013
  51. 79Book2012 IEEE Conference on Computer Vision and Pattern RecognitionD. Ciresan et al. — 2012
  52. 81Very Deep Convolution Networks for Large Scale Image RecognitionKaren Simonyan et al. — 2014
  53. 82JournalGoing deeper with convolutionsChristian Szegedy — 2015
  54. 83Building High-level Features Using Large Scale Unsupervised LearningAndrew Ng et al. — 2012
  55. 84BookDeep LearningIan Goodfellow and Yoshua Bengio and Aaron Courville — MIT Press — 2016
  56. 85BookNonlinear System Identification: NARMAX Methods in the Time, Frequency, and Spatio-Temporal DomainsS. A. Billings — Wiley — 2013
  57. 86Generative Adversarial NetworksIan Goodfellow et al. — 2014
  58. 87A possibility for implementing curiosity and boredom in model-building neural controllersJürgen Schmidhuber — MIT Press/Bradford Books — 1991
  59. 88JournalGenerative Adversarial Networks are Special Cases of Artificial Curiosity (1990) and also Closely Related to Predictability Minimization (1991)Jürgen Schmidhuber — 2020
  60. 90Progressive Growing of GANs for Improved Quality, Stability, and VariationT. Karras et al. — 26 February 2018
  61. 92JournalDeep Unsupervised Learning using Nonequilibrium ThermodynamicsJascha Sohl-Dickstein et al. — PMLR — 1 June 2015
  62. 93Very Deep Convolutional Networks for Large-Scale Image RecognitionKaren Simonyan et al. — 10 April 2015
  63. 94Delving Deep into Rectifiers: Surpassing Human-Level Performance on ImageNet ClassificationKaiming He et al. — 2016
  64. 95Deep Residual Learning for Image RecognitionKaiming He et al. — 10 December 2015
  65. 96Highway NetworksRupesh Kumar Srivastava et al. — 2 May 2015
  66. 97Book2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)Kaiming He et al. — IEEE — 2016
  67. 101Linear Transformers Are Secretly Fast Weight ProgrammersImanol Schlag et al. — Springer — 2021
  68. 102BookProceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System DemonstrationsThomas Wolf et al. — 2020
  69. 103JournalApplication of Artificial Intelligence to the Management of Urological CancerMaysam F. Abbod — 2007
  70. 104JournalAn artificial neural network approach to rainfall-runoff modellingChristian W. Dawson — 1998
  71. 106JournalWhat Is Machine Learning, Artificial Neural Networks and Deep Learning?-Examples of Practical Applications in MedicineJakub Kufel et al. — 2023-08-03
  72. 107JournalA brief review of feed-forward neural networksMurat H. Sazlı — 2006-01-01
  73. 108BookArtificial intelligenceAddison-Wesley Pub. Co — 1992
  74. 109BookSimulation neuronaler NetzeAndreas Zell — Addison-Wesley — 2003
  75. 110JournalA Survey of Partially Connected Neural NetworksD. Elizondo et al. — October 1997
  76. 112Lecture Notes: Neural Network ArchitecturesEvelyn Herberg — 2023-04-18
  77. 113BookSimulation Neuronaler NetzeAndreas Zell — Addison-Wesley — 1994
  78. 115BookThe nature of statistical learning theoryVladimir N. Vapnik et al. — Springer — 1998
  79. 116BookFundamentals of machine learning for predictive data analytics: algorithms, worked examples, and case studiesJohn D. Kelleher et al. — The MIT Press — 2020
  80. 118JournalExtreme learning machine: theory and applicationsGuang-Bin Huang et al. — 2006
  81. 119JournalThe no-prop algorithm: A new learning algorithm for multilayer neural networksBernard Widrow — 2013
  82. 120Training recurrent networks without backtrackingYann Ollivier et al. — 2015
  83. 123BookDeep learningIan Goodfellow et al. — The MIT press — 2016
  84. 124JournalHyperparameter optimization: Foundations, algorithms, best practices, and open challengesBernd Bischl et al. — March 2023
  85. 125Forget the Learning Rate, Decay LossJiakai Wei — 26 April 2019
  86. 126Book2009 International Conference on Computational Intelligence and Natural ComputingY. Li et al. — 1 June 2009
  87. 127BookIntroduction to machine learningEtienne Bernard — Wolfram Media — 2021
  88. 128BookIntroduction to Machine LearningEtienne Bernard — Wolfram Media Inc — 2021
  89. 129JournalMetaheuristic design of feedforward neural networks: A review of two decades of researchVarun Kumar Ojha et al. — 1 April 2017
  90. 130Genetic reinforcement learning for neural networksDominic, S. — IEEE — July 1991
  91. 131JournalProcess control via artificial neural networks and reinforcement learningJ.C. Hoskins — 1992
  92. 132BookNeuro-dynamic programmingD.P. Bertsekas et al. — Athena Scientific — 1996
  93. 133JournalComparing neuro-dynamic programming algorithms for the vehicle routing problem with stochastic demandsNicola Secomandi — 2000
  94. 134Neuro-dynamic programming for the efficient management of reservoir networksde Rigo, D. — Modelling and Simulation Society of Australia and New Zealand — 2001
  95. 135Genetic algorithms and neuro-dynamic programming: application to water supply networksDamas, M. — IEEE — 2000
  96. 136BookOptimization in MedicineGeng Deng — 2008
  97. 138JournalSelf-learning agents: A connectionist theory of emotion based on crossbar value judgmentStevo Bozinovski et al. — 2001
  98. 139Evolution Strategies as a Scalable Alternative to Reinforcement LearningTim Salimans et al. — 7 September 2017
  99. 140Deep Neuroevolution: Genetic Algorithms Are a Competitive Alternative for Training Deep Neural Networks for Reinforcement LearningFelipe Petroski Such et al. — 20 April 2018
  100. 142JournalA Learning Algorithm for Boltzmann MachinesDavid H. Ackley et al. — 1985
  101. 143BookStochastic Models of Neural NetworksClaudio Turchetti — IOS Press — 2004
  102. 144MagazineHands-On Bayesian Neural Networks—A Tutorial for Deep Learning UsersLaurent Valentin Jospin et al. — 2022
  103. 145JournalTopologyNet: Topology based deep convolutional and multi-task neural networks for biomolecular property predictionsZixuan Cang et al. — 27 July 2017
  104. 147BookApplied Soft Computing Technologies: The Challenge of ComplexityC. Ferreira — Springer-Verlag — 2006
  105. 150JournalA learning algorithm of CMAC based on RLSTing Qin et al. — 2004
  106. 151JournalContinuous CMAC-QRLS and its systolic arrayTing Qin et al. — 2005
  107. 152JournalTunability: Importance of Hyperparameters of Machine Learning AlgorithmsPhilipp Probst et al. — 26 February 2018
  108. 153Neural Architecture Search with Reinforcement LearningBarret Zoph et al. — 4 November 2016
  109. 154JournalAuto-keras: An efficient neural architecture search systemHaifeng Jin et al. — ACM — 2019
  110. 155Hyperparameter Search in Machine LearningMarc Claesen et al. — 2015
  111. 156JournalTuring computability with neural netsH.T. Siegelmann et al. — 1991
  112. 157NewsAnalog computer trumps Turing modelSunny Bains — 3 November 1998
  113. 158JournalComputational Power of Neural Networks: A Kolmogorov Complexity CharacterizationJosé Balcázar — July 1997
  114. 159BookProceedings of the 27th ACM International Conference on MultimediaFriedland Gerald — ACM — 2019
  115. 161BookInformation Theory, Inference, and Learning AlgorithmsDavid J.C. MacKay — Cambridge University Press — 2003
  116. 162JournalWide neural networks of any depth evolve as linear models under gradient descentJaehoon Lee et al. — 2020
  117. 164BookNeural Information ProcessingXu ZJ, Zhang Y, Xiao Y — Springer, Cham — 2019
  118. 165JournalOn the Spectral Bias of Neural NetworksNasim Rahaman et al. — 2019
  119. 166Theory of the Frequency Principle for General Deep Neural NetworksTao Luo et al. — 2019
  120. 168BookHandbook of Applied MathematicsRobin Esch — Springer US — 1990
  121. 169BookA Concise Guide to Market ResearchMarko Sarstedt et al. — Springer Berlin Heidelberg — 2019
  122. 170Book2016 IEEE Symposium Series on Computational Intelligence (SSCI)Jie Tian et al. — December 2016
  123. 171BookDynamic Data Assimilation – Beating the UncertaintiesWesam Salah Alaloul et al. — 2019
  124. 172Book2013 International Conference Oriental COCOSDA held jointly with 2013 Conference on Asian Spoken Language Research and Evaluation (O-COCOSDA/CASLRE)Madhab Pal et al. — IEEE — 2013
  125. 174JournalLung sound classification using cepstral-based statistical featuresNandini Sengupta — August 2016
  126. 1753D-R2N2: A Unified Approach for Single and Multi-view 3D Object ReconstructionChristopher B. Choy et al. — 2016
  127. 176JournalIntroduction to Neural Net Machine VisionTurek, Fred D. — March 2007
  128. 177Book2015 13th International Conference on Document Analysis and Recognition (ICDAR)Durjoy S. Maitra et al. — August 2015
  129. 179JournalThe time traveller's CAPMJordan French — 2016
  130. 180JournalNeural network approach to quantum-chemistry data: Accurate prediction of density functional theory energiesRoman M. Balabin et al. — 2009
  131. 183NewsFacebook Boosts A.I. to Block Terrorist PropagandaSam Schechner — 15 June 2017
  132. 184BookIntroduction to Artificial Intelligence: from data analysis to generative AIAlberto Ciaramella et al. — Intellisemantic Editions — 2024
  133. 185Book2026 IEEE Madhya Pradesh Section Conference (MPCON)Abhishek Maity — March 2026
  134. 186JournalApplication of Neural Networks in Diagnosing Cancer Disease Using Demographic DataN Ganesan — 2010
  135. 189JournalChanges in cell shape are correlated with metastatic potential in murineSamanthe Lyons — 2016
  136. 190JournalDeep Learning for Accelerated Reliability Analysis of Infrastructure NetworksMohammad Amin Nabian et al. — 28 August 2017
  137. 192JournalUse of artificial neural networks to predict 3-D elastic settlement of foundations on soils with inclined bedrockE. Díaz et al. — September 2018
  138. 194JournalArtificial Neural Networks in Hydrology. I: Preliminary ConceptsRao S. Govindaraju — 1 April 2000
  139. 195JournalArtificial Neural Networks in Hydrology. II: Hydrologic ApplicationsRao S. Govindaraju — 1 April 2000
  140. 196JournalSignificant wave height record extension by neural networks and reanalysis wind dataD. J. Peres et al. — 1 October 2015
  141. 198JournalArtificial Neural Networks applied to landslide susceptibility assessmentLeonardo Ermini et al. — 1 March 2005
  142. 199Book2017 International Joint Conference on Neural Networks (IJCNN)R. Nix et al. — May 2017
  143. 201BookCyber Threat IntelligenceSajad Homayoun et al. — Springer International Publishing — 2018
  144. 202BookProceedings of the Twenty-Seventh Hawaii International Conference on System Sciences HICSS-94Ghosh et al. — January 1994
  145. 206JournalVariational Quantum Monte Carlo Method with a Neural-Network Ansatz for Open Quantum SystemsAlexandra Nagy — 28 June 2019
  146. 207JournalConstructing neural stationary states for open quantum many-body systemsNobuyuki Yoshioka et al. — 28 June 2019
  147. 208JournalNeural-Network Approach to Dissipative Quantum Many-Body DynamicsMichael J. Hartmann et al. — 28 June 2019
  148. 209JournalSimulation of alcohol action upon a detailed Purkinje neuron model and a simpler surrogate model that runs >400 times fasterForrest MD — April 2015
  149. 211JournalScaling deep learning for materials discoveryAmil Merchant et al. — December 2023
  150. 212JournalAdvances in Artificial Neural Networks – Methodological Development and ApplicationYanbo Huang — 2009
  151. 213JournalExploring the Advancements and Future Research Directions of Artificial Neural Networks: A Text Mining ApproachElham Kariri et al. — 2023
  152. 215JournalDeep Learning for Intelligent Human–Computer InteractionZhihan Lv et al. — 2022-11-11
  153. 217JournalDeep networks for system identification: A surveyGianluigi Pillonetto et al. — 2025-01-01
  154. 218JournalA Survey of Data Mining and Machine Learning Methods for Cyber Security Intrusion DetectionAnna Buczak et al. — 2016
  155. 219JournalGenerative AI and ChatGPT: Applications, challenges, and AI-human collaborationFiona Fui-Hoon Nah et al. — 3 July 2023
  156. 221JournalFrom artificial neural networks to deep learning for music generation: history, concepts and trendsJean-Pierre Briot — January 2021
  157. 222JournalGhost in the (Hollywood) machine: Emergent applications of artificial intelligence in the film industryPei-Sze Chow — 6 July 2020
  158. 223BookThe 3rd International Conference on Information Sciences and Interaction SciencesXinrui Yu et al. — IEEE — June 2010
  159. 224JournalContinual lifelong learning with neural networks: A reviewGerman I. Parisi et al. — 1 May 2019
  160. 225JournalGrowing pains for deep learningChris Edwards — 25 June 2015
  161. 229JournalStatistical process monitoring of artificial neural networksAnna Malinovskaya et al. — January 2024
  162. 230JournalAddressing bias in big data and AI for health care: A call for open scienceNatalia Norori et al. — October 2021
  163. 231JournalFailing at Face Value: The Effect of Biased Facial Recognition Technology on Racial Discrimination in Criminal JusticeWang Carina — 27 October 2022
  164. 234JournalA hybrid neural networks-fuzzy logic-genetic algorithm for grade estimationTahmasebi et al. — 2012