Abstract
The conventional interpretation of spikes is from the perspective of an external observer with knowledge of a neuron’s inputs and outputs who is ignorant of the contents of the “black box” that is the neuron. Here we consider a neuron to be an observer and we interpret spikes from the neuron’s perspective. We propose both a descriptive hypothesis based on physics and logic, and a prescriptive hypothesis based on biological optimality. Our descriptive hypothesis is that a neuron’s membrane excitability is “known” and the amplitude of a future excitatory postsynaptic conductance (EPSG) is “unknown”. Therefore excitability is an expectation of EPSG amplitude and a spike is generated only when EPSG amplitude exceeds its expectation (“prediction error”). Our prescriptive hypothesis is that a diversity of synaptic inputs and voltage-regulated ion channels implement “predictive homeostasis”, working to insure that the expectation is accurate. The homeostatic ideal and optimal expectation would be achieved when an EPSP reaches precisely to spike threshold, so that spike output is exquisitely sensitive to small variations in EPSG input. To an external observer who knows neither EPSG amplitude nor membrane excitability, spikes would appear random if the neuron is making accurate predictions. We review experimental evidence that spike probabilities are indeed maintained near an average of 0.5 under natural conditions, and we suggest that the same principles may also explain why synaptic vesicle release appears to be “stochastic”. Whereas the present hypothesis accords with principles of efficient coding dating back to Barlow (), it contradicts decades of assertions that neural activity is substantially “random” or “noisy”. The apparent randomness is by design, and like many other examples of apparent randomness, it corresponds to the ignorance of external macroscopic observers about the detailed inner workings of a microscopic system.
Introduction
Within the field of information theory, there is a common though often implicit belief that the information content of a signal depends only on its statistical relationship to the quantity of interest (e.g., Rieke et al., ). In this view, we should be able to understand the meaning of a neuron’s spikes entirely through an input-output (I-O) analysis, without any consideration of events happening within the neuron. This view of information is a natural consequence of the opinion that probabilities are essentially equivalent to frequencies and are properties of an observed system, rather than being entirely conditional on the knowledge possessed by an observer about that system (Fiorillo, ). The faults and limitations of the “frequentist” definition of probability and information within neuroscience have been discussed previously, together with the virtues of the “Bayesian” definition advocated by Jaynes (Fiorillo, ). The present work considers a neuron as an observer that knows only what is in its internal biophysical state. Like Jaynes (), we presume that probabilities are always conditional on the local and subjective information of an observer, through the universal and objective principle of logic (Fiorillo, ).
At the center of the present work is a very simple idea that has been around at least since the advent of information theory (Shannon, ). If an event has two possible outcomes, observation of the outcome will convey the maximum amount of information if the probability of each is equal prior to observation. The output of most neurons is a spike that is virtually all-or-none (unlike a graded output, this allows for reliable communication across distances), and thus the amount of information that could be conveyed by a spike will be maximal when its probability is 1/2. Although this is not controversial, it has nonetheless been associated with considerable confusion. This is because the information conveyed by any event will necessarily depend entirely on the prior knowledge of the observer, and a signal that is highly informative to one observer could be entirely uninformative to another. Without a clear understanding of how probabilities and information depend on the observer, scientists have often misattributed their information to the neurons they study (Fiorillo, ).
Consider the case of selecting questions that must have “yes” or “no” answers (as in the game “20 Questions”). If there is a real number known to be within the range of 0–10, and the goal is to guess it as accurately as possible after receiving the answer to a yes-no question, then the best strategy would be to bisect the range of possible numbers evenly by asking “is the number greater than 5?” This is the optimal question because “5” is the expected value and will maximize the questioner’s uncertainty (entropy) by causing both “yes” and “no” to have equal probabilities of 1/2. The answer will therefore provide the maximal amount of information by eliminating this uncertainty. Any other question would result in less initial uncertainty about the answer, and therefore the answer would convey less information. However, the answer is informative only if one knows the question. A second observer may also know that the answer will be “yes” or “no”, but not know the question. Thus two observers could agree with one another that the probability of “yes” is 1/2, but the one who knows the question could get maximal information, whereas the one who is entirely ignorant of the question would learn nothing useful from the same answer, and might be tempted to declare that the answer was entirely random noise. Here we suggest that in many instances neuroscientists have essentially taken the perspective of the naive observer, being ignorant of the prior information of the neuron that “asks the question” and generates and interprets the spikes.
It is well known that the spike output of neurons often appears “stochastic”, insofar as the same stimulus input causes a variable spike output. Likewise, when an action potential arrives at a synaptic terminal, it may or may not causes release of neurotransmitter. Although it has been noted that this variability in I-O relations could convey information (Softky and Koch, ; Softky, ), over decades it has repeatedly been attributed to “noise” or “randomness” that is presumed to degrade the ability of neurons to transmit information (Calvin and Stevens, ; Tolhurst et al., ; Shadlen and Newsome, , ; London et al., ). Models of how neurons could perform inference have proposed that the variability may serve a function by signifying subjective uncertainty (Pouget et al., ; Deneve et al., ; Hoyer and Hyvarinen, ; Jazayeri and Movshon, ; Ma et al., ; Beck et al., ; Berkes et al., ). Whether the variability is noise or a functional signal related to uncertainty, uncertainty (ignorance) itself is not desirable, and both of these viewpoints propose that a more variable output (in response to a fixed input) would correspond to greater uncertainty. In contrast, a central conclusion of the present work is that this same variability is a signature of optimality and knowledge if we consider the perspective of a neuron, which performs work (consumes energy) to produce it.
It may seem strange or even unscientific to treat neurons as observers. Indeed, Skinner and other “behaviorists” argued that to be objective and to give psychology a firm scientific basis, we should not even treat humans and other animals as observers, but that we should instead characterize their behavior solely in terms of I-O relations (environmental inputs and behavioral outputs). Psychology has now mostly rejected that view, and instead seeks to understand mental events and behavior from the first-person perspective of the brain as it observes its world. The probability theory of Jaynes () allows us to describe a subjective state of knowledge in an objective manner (Fiorillo, ). Many elegant studies have now demonstrated that aspects of perception, cognition, and motor control that appeared maladaptive can be understand as optimal (rational or Bayesian) if we take the brain’s perspective, including sensory illusions (e.g., Weiss et al., ; Niemeier et al., ; Yang and Purves, ; Körding and Wolpert, ).
Illusions would seem to provide clear evidence of the brain’s malfunction. However, the brain only appears to function poorly from the perspective of a third-party observer who knows something that the brain itself does not and could not know. Whereas the brain must infer the sensory stimulus given limited sensory evidence and its own prior knowledge of what is likely, the third party has privileged knowledge of the “true sensory stimulus”. The brain integrates its information as well as it possibly can, but it nonetheless “misperceives” the stimulus on those rare occasions when the stimulus is something unusual and unexpected, like a fast moving object (Weiss et al., ). Thus, what once appeared pathological from our perspective as external observers can suddenly be understood as optimal once we consider the unique perspective of the neural observer. We propose that, in a precisely analogous manner, the variability in a neuron’s I-O relationship is optimal rather than pathological once one takes the neuron’s perspective.
Information theory from the neuron’s perspective
Our theory is closely related to Barlow’s theory of “efficient coding” (Barlow, ) and the extensive literature that followed on the relation between neural activity and natural stimulus statistics (e.g., Laughlin, ; Srinivasan et al., ; Linsker, ; Rieke et al., ; Stemmler and Koch, ; Brenner et al., ; Simoncelli and Olshausen, ; Hosoya et al., ; Sharpee et al., ). In particular, the frequency distribution of a neuron’s output (e.g., firing rate) should be “matched” to the distribution of its input intensity, so that if the input distribution is Gaussian, then the I-O relation should ideally follow the form of the corresponding cumulative Gaussian (Laughlin, ; Brenner et al., ). A neuron’s membrane excitability determines its I-O relationship, and excitability should adapt so that the location (x-offset) and scale (slope) of the I-O curve cause the neuron’s output to be most sensitive to the most probable inputs (the steep and linear portion of the curve) at the expense of improbable inputs (the non-linear and nearly flat portion of the curve). Figure 1 illustrates this principle as it applies to the cellular level, with the amplitude of excitatory postsynaptic conductance (EPSG) as input and a binary spike or no spike as output. We refer to the principle as “predictive homeostasis”, rather than “predictive” or “efficient” coding, to emphasize that it is implemented by numerous homeostatic mechanisms of the same sort that are well known by biologists to be present in all cells.
Figure 1
Efficient coding by a neuron should ideally have the effect of maximizing the Shannon information in its spike output (IS) about its excitatory synaptic input (EEPSG). This information would be equal to the reduction in entropy (H) caused by observation of the spike output (S).
The “prior entropy” [H(EEPSG|EX)] would be the width or uncertainty of the probability distribution, which would be reduced by observing spike output (S). We equate a neuron’s membrane excitability with the prior information (EX). Unfortunately we are unable at this time to derive the prior probability distribution [p(EEPSG|EX)], as discussed previously (Fiorillo, ). However, since spike output (S) is binary, its information content is maximized when its prior probability [p(S|EX)] is 1/2. This corresponds to that case that it indicates whether the input is more or less than the prior expectation (<EEPSG>|EX), so that observation of spike output reduces the prior entropy [H(EEPSG|EX)] by 1/2 (assuming the prior distribution is symmetric). We propose below that this is precisely the meaning of a spike from the neuron’s point of view.
There are at least three respects in which the present hypothesis goes beyond previous conceptions of this principle. First and of greatest importance, we consider the neuron’s point of view, by which we mean that the expectation of interest is entirely conditional on the neuron’s prior information as inherent in its membrane excitability (EX). This contrasts with the conventional approach in which probabilities are conditional on the knowledge of an external observer (e.g., Rieke et al., ). Although the importance of “taking the neuron’s point of view” has been recognized previously (Rieke et al., ), it has virtually never been done for reasons described elsewhere (Fiorillo, ). It requires that we set aside whatever knowledge we might have about the neuron’s input (such as knowledge of the frequency distribution of its inputs, commonly referred to as “the stimulus ensemble”) and its statistical relationship to the neuron’s output. Instead, we consider only the information that we believe to be contained within the physical structure of the neuron or other observer.
A second distinction from most previous work is in considering the millisecond timescale, in which spike output is necessarily binary. Most previous work on efficient coding, including our own (Fiorillo, ), considered output to be analog firing rate or membrane voltage (Laughlin, ; Linsker, ; Stemmler and Koch, ; Brenner et al., ; but see Deneve, ). However, all the relevant membrane properties evolve on the scale of 1 ms or even less (Softky, ), including synaptic conductance and the action potential itself. There are many slower processes as well, some of which effectively average over spikes, and these are also incorporated within the theory (Fiorillo, ). But to adequately capture the informational dynamics, and especially the proximal cause of spikes, we must consider the fastest time scale in which membrane properties vary. Our approach rests on the intuitive notion that information is inseparable from matter, energy, and causation, so that if the charge distribution across a neuron’s membrane is changing within a millisecond, so is the information that it contains.
Third, like Stemmler and Koch (), we focus on a single neuron at the cellular level, so that input and output are both within the neuron. This contrasts with most previous work on efficient coding in which input corresponds to external sensory stimuli, and thus multiple neurons intervene between input and spike output. The cellular approach has several advantages. First, as a simpler system, analysis of a single neuron better allows us to identify mechanisms by which neurons predict their input and adapt their I-O relation (e.g., Hong et al., ). Second, whereas efficient coding is often thought of as a “sensory problem”, the cellular analysis emphasizes the general nature of the problem, since all spiking neurons are faced with the problem of using a binary output to efficiently represent an analog input. Third, the inference made by a neuron or any observer must be based entirely on local information, as must all learning. Compared to a network of neurons, it is relatively easy to imagine how learning or natural selection might optimize I-O relations of single neurons (Fiorillo, ). If a single neuron can optimize its local I-O relation, a feedforward series of such neurons, each adapted to the local statistics of its input, can be expected to approximately optimize the systems-level I-O relation (between external sensory stimuli and spike output).
A previous model of prediction by a single neuron
A previous theory had the goal of providing a general informational account of the nervous system (Fiorillo, ). Although narrower in scope, the present work builds on the previous framework. In both cases, a neuron’s membrane compares the intensity of its sensory related input to its expectation, and its output signals “prediction error”. As in other models of neuronal prediction error (e.g., Schultz et al., ; Izhikevich, ; Frémaux et al., ), the homeostatic null point was presumed to correspond to “baseline firing rate” or an intermediate membrane voltage “just below” spike threshold (Fiorillo, ). The null point and its neural basis was not precisely defined or justified. Indeed, the notion of a “firing rate” is itself problematic, since it is not a physical entity that exists at any particular point in space and time. The present work provides a solution by considering the timescale of a single spike and proposing that the threshold voltage for a spike corresponds to the homeostatic point at which the prediction error is 0.
Information necessarily has both a qualitative and quantitative component, and the earlier theory dealt with both (Fiorillo, ). Like Shannon (), we deal here only with the quantitative aspect. Since EPSGs are the proximal cause of spikes, the set of excitatory synapses will determine the qualitative nature of the information within spikes by specifying the neuron’s receptive field, or “stimulus” (for example, light of certain wavelengths in some region of space). Previous work suggested how learning could shape a receptive field to insure that it contributes information of biological importance (e.g., Schultz et al., ; Izhikevich, ; Fiorillo, ; Frémaux et al., ). Here we address only the issue of how membrane excitability can enable spike output to best preserve (maximize) information about EPSG input (the “coding strategy”), regardless of the meaning and importance of that input (and “what it should code”).
Theory of predictive homeostasis
After specifying a simple biophysical model of membrane excitability, we proceed to describe its information content based only on its physical properties (Hypothesis 1), and then to prescribe excitability based on biological optimality (Hypothesis 2).
The biophysical model
We assume a generic neuron that follows established biophysics. It has a spike generating mechanism as well as two conductances, an EPSG and an “intrinsic” conductance. In general, each of these will vary over time, and will be the sum of multiple individual conductances (synapses or types of ion channel). EPSGs are discrete events consisting of one or more unitary EPSGs, where a unitary EPSG results from a spike in a single presynaptic neuron (which releases at least one vesicle of neurotransmitter, and usually more). Unitary EPSGs vary in amplitude over time, even in the case that they derive from the same presynaptic neuron (due to variable numbers of vesicles being released across multiple release sites). In addition, unitary EPSGs and their associated EPSPs will display spatial and temporal summation. Although, they are discrete events and may occur only sparsely over time, they convey information about continuously varying external variables, such as light or sound intensity. We focus on the amplitude of EPSGs because although they are internal to a neuron, their timing and amplitude depend almost entirely on variables external to the cell.
Whether an EPSG arriving at a synapse causes a spike to be initiated near the soma depends on many well known factors, including initial membrane voltage, axial resistance, and membrane capacitance and conductance. Voltage-regulated ion channels change membrane conductance to amplify or suppress an EPSP, even during its brief rise time (∼1 ms). The summed effect of all of these factors is encapsulated by the term “excitability”. We define it here as the energy barrier (EX) which separates an EPSG from spike threshold, although to be consistent with common usage we would say that a lower energy barrier (a smaller distance from spike threshold) corresponds to a higher excitability. A spike will occur only if the energy in an EPSG (EEPSG) exceeds the barrier (EX). Thus excitability is the amount of work that an EPSG must perform to cause a spike, and this corresponds to the amount of charge that must be transferred from the synapse to the site of spike initiation. This could be calculated in principle given a detailed model, but it can be found more easily by injecting “test EPSGs” of incrementally increasing amplitude until spike threshold is reached (Figure 2A). The amplitude of the “threshold EPSG” that is found in this way has precisely the energy EX that it is needed to reach spike threshold.
Figure 2
Describing a neuron’s excitability and expectation
Here we equate “the observer” with “membrane excitability”. The excitability at time t will determine whether an EPSG with onset at time t will cause a spike. What is the expectation of EPSG amplitude given only excitability? Some would answer that this question cannot be answered unless one first observes EPSG amplitudes and thus has knowledge of the frequency of various amplitudes. However, we follow the probability theory of Jaynes (commonly referred to as “Bayesian”), according to which one “sample” of any quantity can provide an expectation of another (through logic, as expressed in the principle of maximum entropy) (Jaynes, ; Fiorillo, ). From this we presume that knowledge of one mass can be used to estimate another mass, one energy can be used to estimate another energy, etc. More information is always better for a biological observer, but an observer simply “knows what it knows”.
The EPSG performs work to drive membrane voltage towards spike threshold, and excitability works against the EPSG. In the absence of any information beyond excitability itself, the probability that an EPSG will cause a spike is 1/2 (based on logic, the maximum entropy principle), and thus the expectation (<EEPSG>) must be equal to the excitability (EX) (equation 2) (by the mathematical definition of “expectation”).
Hypothesis 1: A spike is generated only if EPSG amplitude is greater than expected given the information in membrane excitability (equation 2).
Equation 2 refers to the expectation (at time t) of the peak amplitude of an EPSG if one were to start at that same time (which would reach its peak at time t + r, where r refers to the stereotyped rise time of an EPSG, typically about 0.5 ms). Thus, we are referring to an expectation of a potential future event. When excitability is low, the energy barrier (EX) is high, and a large EPSG is expected. The expectation is just the “center” of a probability distribution, and we are not speculating here about the uncertainty (width) of the distribution. Because spike threshold corresponds to the expectation, a spike conveys maximal information (relative to any other relation between expectation and spike threshold, assuming that the prior probability distribution is symmetrical). A spike would not convey any information beyond that in membrane voltage at the spike initiation site at the time of its generation, but it would reliably communicate that information throughout the entire neuron.
Given an EPSG with onset at time t, it is useful to denote membrane excitability at time t as “prior information” to distinguish it from the new information in the EPSG. Hypothesis 1 is essentially just that a spike signals “prediction error”. Prediction errors are known to be efficient and useful signals, but there is not much intelligence in a prediction error if the prediction itself is not accurate. Our use of “prediction error” is merely descriptive and could be applicable to a large variety of physical entities.
A traditional balance scale provides a useful analogy. It consists of an arm that rotates around a central joint depending on a known reference weight on the left (excitability) and an unknown weight on the right (the EPSG). The arm rotates continuously (over some range) as a function of the difference between the two weights (membrane voltage). The scale generates a binary output to signal which weight is greater (right side up or down, analogous to a spike). Prior to placement of the unknown weight on the right side, the expectation of the unknown weight would be equal to the known weight, and “right side up” (a spike) would indicate that the unknown weight was greater than the expectation.
Our challenge here is merely to describe the information in a physical entity based on physics and logic. Unlike most of biology, physics is purely descriptive insofar as it is agnostic to “how the world should work”. In describing the knowledge within a balance scale or a neuron, we do not presume that the reference weight or the excitability is well suited for measuring the unknown weight or EPSG. Rather we presume that it is the only information possessed by the observer.
Critique of hypothesis 1
We believe that an observer is something remarkably simple. Excitability is simple insofar as it is just a single number at a single moment in time. We presume that knowledge of excitability does not include knowledge of its constituents, not even membrane voltage and conductance. Furthermore, excitability is entirely internal to a neuron, and therefore it could be said that we are “taking the neuron’s point of view”. However, we also believe that a single observer should correspond to information that is physically integrated in space and time. Membrane voltage fits this definition, summing currents over a local region of membrane. By contrast, excitability includes multiple factors in addition to membrane voltage. Even though those factors are all present in a small and well defined part of a neuron, they are not being physically summed together. Therefore we infer that excitability corresponds to information distributed across more than one observer, and Hypothesis 1 falls short of our ideal of “describing the knowledge of an observer”. It remains for future work to describe the knowledge in membrane voltage about EPSG input.
Despite its limitations, we conclude that Hypothesis 1 represents a useful default hypothesis by virtue of its simplicity. For the expectation of EPSG amplitude given excitability to be other than excitability itself, and the prior probability of a spike to be other than 0.5, would require additional information that we have not assumed. We assume that our observer has knowledge, but it is an absolute minimal level of knowledge. The prescriptive hypothesis presented below does not depend upon the veracity of this descriptive hypothesis. However, the simple logic of this descriptive hypothesis suggests that a neuron does not need any particular design in order for it to have an expectation, and for its spikes to be maximally informative from its perspective. Hypothesis 1 states that a known is the best guess of an unknown, whereas Hypothesis 2 states that biology works towards the ideal in which the known is in fact equal to the unknown.
Prescribing a neurons excitability and expectation
Hypothesis 2 expresses our belief that the excitability of a healthy adult neuron will tend to accurately predict EPSG amplitude, because it has been shaped by both natural selection over generations and associative learning rules. An accurate analog expectation of an analog variable will be too high in half of cases, and too low in the other half.
Hypothesis 2: Expectations will be accurate under natural conditions, and therefore half of EPSGs will cause spikes.
The perfect match of EPSG amplitude to its expectation is not attainable, and even if it were, the spike output must be 0 or 1, corresponding to a negative or positive prediction error. The optimal expectation would therefore result in a spike in response to precisely half of EPSGs (Figure 1). The spike probability (SP) of 1/2 given the neuron’s prior information (Equation 2) would then match the frequency distribution of spikes given EPSGs. Although this optimal neuron would have low subjective uncertainty (which we do not attempt to quantify here), its spikes would appear to be random to an external observer who is ignorant of variations in excitability and EPSG amplitude.
Hypothesis 1 proposes that spikes are maximally informative simply as a consequence of logic and physics, regardless of the prior information in excitability. Hypothesis 2 implies that accurate prediction (or minimizing total entropy) is an important biological goal, and it will be aided by accurate prior information (a low prior entropy, [H(EEPSG|EX)]. As a contrived example, neuron “A” has prior knowledge that EPSG amplitude must be between 0 and 10, and neuron “B” knows that it must be between 4 and 6. If the expectation of each neuron is 5, then a spike would tell neuron “A” that the actual amplitude must be between 5 and 10, whereas it would indicate to neuron “B” that it is between 5 and 6. Neuron “B” would have more total information given its prior information and a spike [less “posterior” entropy, H(EEPSG|EXS)].
Measuring the accuracy of a neuron’s predictions
Here we propose methods by which we can measure the accuracy of a neuron’s expectation. The neuron itself also needs to assess the accuracy of its predictions, but we address this in a later section (Prediction Errors in Learning). We use “accuracy” to refer to a distance between two expectations, the more certain of which could be called “reality” or the “actual” or “true” value, and the less certain “the expectation”. Thus “accuracy” is not the same as uncertainty (entropy). We have not attempted to quantify uncertainty, but greater uncertainty would naturally tend to be associated with less accuracy.
Residuals
The best method for us to assess the accuracy of expectations will depend on our ability to know and control the neuron. If we have complete knowledge and control, as in a Hodgkin-Huxley model, then it is easy to measure excitability by injecting “test EPSGs” of varying amplitude to find the “threshold EPSG” amplitude that causes an EPSP to reach precisely to spike threshold (Figure 2A). This “threshold EPSG” is our measure of excitability (EX) and corresponds to the neuron’s expectation of EPSG amplitude (Hypothesis 1). The “residual” EPSG amplitude (Eres) measures the difference between the neuron’s expectation (EX) and the actual EPSG (EEPSG) (Figure 2A).
Indices for time have been omitted for simplicity. In principle, excitability could be measured at each instant (“real time”). However, it is not an expectation of EPSG likelihood, but of EPSG amplitude if one were to occur. Thus the expectation is only put to the test when an EPSG occurs. A high level of inhibitory conductance would correspond to expectation of a large EPSG, but if none occurs, the cost would be only metabolic and not informational. Thus we measure excitability only at the time of real EPSGs (onset of test and real EPSGs are always synchronous) (Figure 2A).
A residual (Eres) will be positive if the EPSG causes a spike, and will indicate the extent to which the EPSG was greater than expected and excitability was too high (EX was too low). It will be negative if EPSG amplitude was less then expected, indicating that excitability was too low (EX was too high). If a neuron is making accurate predictions, the residuals will be small with an average of 0, and spike generation will be highly sensitive to small variations in EPSG amplitude, as experiments have demonstrated in some cases (Carandini, ; Kuenzel et al., ).
The residual is an error of sorts, and an objective of learning would be to modify homeostatic conductances so that residuals are minimized. However, we are not suggesting that a neuron has any direct means of knowing the residuals, and we reserve the term “prediction error” to describe errors from the neuron’s point of view (as signaled by spikes). Thus the residuals cannot drive learning, and we present them only as a means for us to assess the accuracy of a neuron’s predictions (its distance from optimality) (Figure 2).
Spike probability from the experimenter’s point of view
Hypothesis 2 states that predictions should be accurate under natural conditions in a normally functioning brain. Even with in vivo intracellular recording (Figure 3) it is challenging to estimate EPSG amplitudes and excitability. It is therefore useful to assess the accuracy of a neuron’s expectations given only knowledge of the time of EPSG inputs (or presynaptic spikes) and spike outputs. For this purpose we define spike probability (SP):
Figure 3
Spike probability (SP): The probability that at least one output spike is caused by an EPSG at a “driving” synapse.
SP is defined to incorporate knowledge that at least one input spike and associated EPSG has occurred at a “driving” synapse (see Driving EPSGs as the Cause of Spikes), and that if the EPSG causes a spike, it will occur within a certain delay (generally less than 5 ms in a typical neuron). A unitary EPSG could be the sole cause of a spike or a partial cause together with other EPSGs that are nearly synchronous. SP does not include knowledge of excitability or the specific amplitude of the EPSG (since that information is seldom available from experimental data). Thus, our uncertainty about whether or not a spike will occur is not because there is anything intrinsically “random” about the occurrence of spikes, but because we are ignorant of the “black box” that is EPSG amplitude and membrane excitability.
SP of 1/2 is the homeostatic ideal and will cause spike output to be maximally sensitive to small gradations in EPSG amplitude (Figure 1). However, we presume that it is a virtually impossible goal to achieve, for the same reason that predictions of the future market value of stocks are almost never precisely accurate. We expect that SP varies dynamically from near 0 to 1 under natural conditions, but with an average over time that is near 1/2.
Implementation of predictive homeostasis
Driving EPSGs as the cause of spikes
We presume that information follows causation, and that the EPSG is (and should be) the proximal cause of spikes. The neuron’s knowledge is substantially limited to its proximal causes, for the same reason that our conscious perception corresponds to proximal sensory causes in the immediate past, which we can infer with relatively little uncertainty, but not the innumerable distal causes further in the past. In a typical neuron the proximal cause of a spike would be an EPSG with onset about 0.5–5 ms earlier, mediated primarily by AMPA-type glutamate receptors (Figure 3). We propose that slow membrane depolarization would not be the proximal cause of spikes under natural conditions, including that initiated by offset of synaptic inhibition, activation of G-protein-coupled receptors, and perhaps even AMPA-type glutamate receptors located on distal dendrites (Branco and Häusser,
Mechanisms and timescales
We propose that there are a vast array of homeostatic ion channels and synapses functioning simultaneously on numerous timescales within a single neuron. Each would make a prediction based on a particular recurring pattern of synaptic excitation or inhibition, like those in Figure 3. Homeostatic synaptic inhibition would be dedicated to spatial patterns of activity across neurons (as in the case of “surround inhibition”), whereas voltage-regulated ion channels would be dedicated to temporal patterns (Fiorillo,
Figure 1 illustrates the case of a neuron that receives EPSGs of only three amplitudes (1, 2, and 3) that are sufficiently separated in time so that there is no temporal summation. Knowing those three amplitudes but nothing else, the optimal expectation is “2” and an EPSG of amplitude “2” should cause an EPSP that reaches precisely to spike threshold (Figure 1). Figure 2 illustrates the case of a neuron that receives excitatory drive from a single presynaptic neuron, with each presynaptic spike causing a unitary EPSG of identical amplitude. Despite the simplification of identical unitary amplitudes, this example remains interesting since temporal summation will cause EPSG and EPSP amplitude to vary with presynaptic inter-spike interval (ISI). This is in fact similar to some real synapses, as discussed further below.
The neuron of Figure 2 is the simplest possible insofar as its only homeostatic conductance is a “leak” conductance, which is voltage-insensitive and constant by definition. If the leak conductance is optimal, it will accurately predict the long-term average EPSG and EPSP amplitude, which vary due to temporal summation. Thus half of unitary EPSGs will exceed average amplitude and cause a spike, and this will only occur when there is sufficient temporal summation (Figures 2A, B). When the neuron’s excitability is “at rest”, the first EPSG in a pair will be of less than expected amplitude and cause no spike, but it will increase excitability so that a second EPSG 5 ms later will cause a spike. We note that even when EPSGs cease and excitability approaches a stable resting level, excitability still provides an expectation of EPSG amplitude, and this could be called the neuron’s “default prior”. Since the amplitude of EPSGs provides information about the intensity of external stimuli, like light intensity, the neuron would always have an expectation of both EPSGs and external stimuli.
Although accurate prediction of the long-term average is important, it would generally be highly inaccurate over shorter periods. Prediction using only a very slow ion channel, like a leak conductance, is analogous to predicting the market value of a stock using only knowledge of its trajectory over the last 10 or 100 years. Better predictions can be made given additional knowledge of any regular pattern, including short term trends. There are a remarkable diversity of voltage-regulated ion channels with a spectrum of kinetic properties, and each of these is proposed to be specialized for making a prediction based on a particular recurring pattern (Fiorillo,
The rising phases of EPSPs and action potentials are the fastest changes in excitatory drive. Excitability will provide an expectation of an EPSG’s amplitude prior to its onset, but A-type K+ channels activate during the rising phase of an EPSP and thus contribute to excitability as well (in many types of neurons). In principle, the amplitude of an EPSP can be predicted by the slope of its rising phase. If some subtypes of A-type channels inactivate within a millisecond or two (Migliore et al.,
Finding the optimal homeostatic conductance
Figure 2 illustrates how we can find the optimal leak conductance for a neuron that has no other homeostatic conductances and receives two EPSGs of equal amplitude that are separated by a 5 ms interval. Because of temporal summation (Figures 2A, B), it is not possible with only a leak conductance to have two EPSPs with both peaks at spike threshold. The best that can be done is for the first to be too small and the second too large (Figure 2C). We presume that the optimal leak conductance is the one that minimizes the sum of the two squared residuals (Figure 2D). If the EPSG interval is 10 rather than 5 ms, there is less temporal summation and thus the optimal leak conductance is smaller (Figure 2D). The optimal conductance in a real neuron would depend on the entire ensemble of EPSG amplitudes that the neuron experiences, and a conductance that is optimal for one pattern will often be suboptimal for another. It is apparent from Figure 2 that if inhibition were increased by an appropriate amount following the first EPSG it would help to restore excitability and thereby make the neuron better prepared for addition of the second EPSG. A neuron with a suitable dynamic homeostatic conductance (like GABAA or A-type channels) would therefore be closer to optimal than any neuron with only a leak conductance.
Two classes of ion channel
Our earlier work distinguished two classes of ion channel (Fiorillo,
The present work provides further insight by specifying that the homeostatic level is spike threshold. GABAA-mediated IPSGs can be class 1 or 2, as illustrated by the two types of GABA synapse on thalamocortical neurons of LGN, both of which derive from local interneurons (Blitz and Regehr,
Sodium channels may also be class 1 or 2. Axonal action potentials are the “all-or-none” output of a neuron, and the sodium channels that cause them clearly drive voltage away from homeostasis and would be class 1. By contrast, it is generally believed that dendritic sodium channels do not consistently contribute to causing axonal action potentials (axonal spikes are not initiated in dendrites). If they usually amplify dendritic EPSPs towards threshold for an axonal spike under natural conditions, they would be homeostatic, whereas if they usually drive depolarization beyond threshold for an axonal spike, they would oppose homeostasis and be in class 1. It is interesting to note that class 1 and class 2 sodium channels could have identical structure at the level of individual channels, but their opposite relationship to homeostasis could be due solely to their sub-cellular localization and density.
Sensory versus motor neurons
The principle of predictive homeostasis should apply to all neurons, and even all cells. However, we suspect that there is an important difference between sensory and motor neurons, and that above we have described sensory neurons. We do not try to justify the distinction here, but a rough and intuitive summary is that sensory neurons use a binary spike output to efficiently convey information about “shades of gray” in their sensory input (EPSG amplitude), whereas motor neurons use it to convert a shade of gray into a “black or white decision”. As partially summarized elsewhere (Fiorillo,
Implications of predictive homeostasis
Metabolic cost of predictive homeostasis and spikes
The importance of predictive homeostasis is evident in the large amount of energy that it consumes. Sodium, potassium, and chloride conductances all tend to increase in synchrony, allowing these ions to flow down their electrochemical gradients and counteract one another electrically while causing only modest changes in membrane voltage (Pouille and Scanziani,
In proposing that SP of 0.5 is optimal, we are ignoring the metabolic cost of action potentials (Attwell and Laughlin,
The meaning of spike output to sub-cellular observers
A spike exerts its effect on sub-cellular observers (e.g., proteins) throughout the neuron. A spike can be approximated as a digital or binary signal, but it acts on virtually all neural observers as an analog signal, perhaps most notably in its effects on calcium concentration. If we were to analyze what these observers know about EPSG amplitude and what information is provided by the spike, we would follow the same framework given above. By predicting some local quantity like calcium concentration, they would also be predicting the more distal EPSG amplitude, which in turn corresponds to information about more distal quantities including a stimulus external to the nervous system.
Will the optimal “encoding strategy” of our neuron benefit these other observers, even though they are ignorant of the membrane excitability? The answer is “yes” because the more effective the predictive homeostasis, the less patterned the spike output, which would be “temporally decorrelated” relative to synaptic input (e.g., Wang et al.,
The benefit of maximizing the variance of input or output can be a source of confusion. It may seem that if the goal of an observer is to minimize its uncertainty (posterior entropy), it would be best to select inputs that have minimum variance and are thus more predictable. This issue has been called “the dark room problem”, since the uncertainty of neurons in the visual system would appear to be minimized if the animal simply avoids light (Friston et al.,
We favor the view that the flow of information follows the flow of force and energy and causation, implying that the absence of one event (a spike) is never the cause of another event, and therefore conveys no information. However, in the context of a force and energy that creates an expectation (e.g., a potassium current), the absence of an event (e.g., an EPSG) could allow the expectation itself to modify the observer’s knowledge (e.g., hyperpolarize the membrane). The occurrence of an EPSP would be known to some dendritic proteins, and the EPSP would cause the expectation of a spike. The absence of a spike in the context of an EPSP would indicate that the EPSP was smaller than expected. By contrast, an axon terminal (or its postsynaptic targets) does not experience an EPSP. Such an observer would have no means to distinguish whether the absence of a spike corresponds to no EPSP or an EPSP of less than expected amplitude, and thus absence of a spike would convey little or no information.
Opponent (ON–OFF) representations
If the absence of a spike conveys little or no information to an axon terminal and its postsynaptic targets, and yet a system is to be sensitive to decrements as well as increments, there must be “opponent” neurons that represent the same stimulus dimension (like light intensity) but with opposite polarity. The ON and OFF neurons in the visual system are a well known example (Wang et al.,
Prediction errors in learning
A spike functions as a “teaching signal”, and what the “student observer” learns will depend on whether or not that observer knows that an EPSG has occurred. Observers within a dendritic spine can and should have information about the occurrence and timing of three events: a local EPSG at that synapse, a global EPSP, and a spike. Given this information and the Hebbian objective at an excitatory synapse of maximizing positive prediction errors (spikes) (Fiorillo,
Whereas observers in the dendrite can gain information from a “spike failure” (EPSG + no spike), downstream observers in the axon terminal or its postsynaptic target neuron would not. For example, midbrain dopamine neurons signal a “reward prediction error” that is thought to teach reward value to downstream neurons (e.g., Schultz et al.,
Comparison to other interpretations of spike variability
Whereas virtually all previous models have proposed that spike variability is indicative of greater uncertainty of neurons about their sensory inputs, we propose the opposite. Our optimal neuron fulfills Softky’s definition of an “efficient neuron”, having membrane excitability that is exquisitely fine-tuned to its synaptic input (Softky,
The characterization of spike variability depends on probabilities, and we have previously criticized the field in general for failing to adequately distinguish conditional probabilities from event frequencies, and for failing to take the neuron’s point of view (Fiorillo,
We view our work, and the EC literature in general, as very distinct from a viewpoint that might be called “probabilistic computation” (PC), which is characterized by the assumption that neurons need to “represent and compute probabilities”. We agree with this viewpoint with respect to the general premise that the function of neurons is related to “Bayesian inference”, but we have a very different view of what this means (Fiorillo,
A particularly comparable example of PC is Deneve’s “Bayesian spiking neuron”, which signals prediction error (Deneve,
Experimental evidence for predictive homeostasis
The theory predicts that homeostatic conductances should counterbalance EPSGs to maintain excitability, and that this should happen on all timescales, including the millisecond scale. The predictions of the theory can be made quantitatively precise given a particular EPSG pattern (Figure 2), as discussed above. In support of the theory, a remarkably precise balance of excitation and inhibition has been found in brain slices (Pouille and Scanziani,
Temporal decorrelation and spike statistics
If a neuron makes accurate predictions and signals prediction error, its spike output should have less temporal pattern than its input, since the predicted patterns have been removed from the output. Multiple studies have demonstrated “temporal decorrelation” within sensory systems (e.g., Dan et al.,
Thalamocortical neurons in LGN
Thalamocortical neurons in LGN offer several advantages for testing the theory. First, as part of the early visual system, they have been studied extensively both in vivo and in vitro, with respect to information theory as well as mechanisms. Second, they have a spiking input and output, like most neurons in the brain but distinct from some early sensory neurons such those in the retina. Third, they have one dominant presynaptic excitatory “driver” that generates large unitary retinogeniculate EPSPs (Usrey et al.,
Table 1 and Figure 4 show output and input rates from studies in LGN and elsewhere in which both were recorded simultaneously. Most studies reported only averages of SP over periods of seconds or minutes. We sometimes report averages of SP across populations of cells, but this is done only to provide a concise summary, or because greater detail was not reported in some studies.
Table 1
| Reference | Cells | Stimulus | N | Input (Hz) | Output (Hz) | O/I(%) |
|---|---|---|---|---|---|---|
| 1. Kaplan et al. ( | Cat LGN | Contrast 64-80% | 37 | NA | NA | 51±20, 19-91 |
| Monkey LGN | Contrast 64-80% | 26 | NA | NA | 71±17, 32-99 | |
| 2. Movshon et al. ( | Monkey LGN | Contrast | 7 | NA | NA | 11-45 |
| 3. Sincich et al. ( | Monkey LGN | Contrast | 15 | 29±9 (n = 13) | 13* | 46±16 |
| 4. Casti et al. ( | Cat LGN | Contrast | 14 | NA | NA | 35±19, 7-70 |
| 5. Weyand (2007) | Cat LGN | Contrast | 12 | 50±16 | 24±8 | 50±13 |
| 6. Lorteije et al. ( | Mouse MNTB | Spontaneous | 11 | 71±11, 0.4-174 | 62* | 87, 32-99 |
| Monaural tone, 80 dB | 15 | 352±34 | 289* | 82, 27-99 | ||
| 7. Hermann et al. ( | Gerbil MNTB | 300 Hz Elec. Stim. | 18 | 300 | 150* | 50* |
| 8. Kopp-Scheinpflug et al. ( | Gerbil MNTB | Spontaneous One tone | 54 | 49 [17-79] 1-308 355 [212-486] 48-868 | 21 [6-45] 0-189 194 [126-230] 18-334 | 48 [36-71] 5-94 55 [38-74 23-95 |
| Two tone | NA | NA | 26 [23-40] | |||
| 9. McLaughlin et al. ( | Cat MNTB | Binaural tone | 49 | 100-500* | 100-500* | 100 |
| 10. Englitz et al. ( | Gerbil MNTB | Spontaneous | 3 | NA | NA | 79, 39, 43 |
| Gerbil AVCN | Spontaneous | 27 | NA | NA | 76±17 | |
| 1 & 2 tone monaural | 34 | NA | NA | 52±19 | ||
| 11. Kuenzel et al. ( | Gerbil AVCN | Spontaneous | 39 | 79* | 56±23 | 71±26 |
| Monaural Tone, 51±11 dB | 39 | NA | NA | 50* | ||
| 12. Chadderton et al. ( | Rat cerebellar granule cells | Tactile, Ipsilateral | 5 | 3.2* EPSCs | 2.2* spikes | 71, 23-150* |
| 13. Bratton et al. ( | Rat lumbar SG | Spontaneous | 38 | 5.4 | 2.9±0.3 | 54* |
| 14. McAllen et al. ( | Rat cardiac SG | Spontaneous | 6 | 4.1±0.7 | 0.2-4.3 | 55* |
Time averaged input and output rates and ratios from studies that measured both simultaneously over seconds or more in cells known to receive driving excitation from just one or a few presynaptic neurons.
Portions of these data are illustrated graphically in Figure 4. The output to input ratio (O/I) should match average SP in principle, but could differ due to a technical failure to detect an input, especially in cases of high frequency bursts of input. Only “12” reported a single cell in which output exceeded input, but “12” was also the only study in which input and output counts were not measured simultaneously. None of these studies reported a significant incidence of either spikes in the absence of input, or more than one output spike resulting from a single input spike. Data are number of neurons (N), population mean±sd (median and quantiles for “8”) and range, except for 3 single cell’s in MNTB for “10.” Asterisks (*) indicate our estimates based either on our calculations or visual inspection of figures. NA: not available. “Spontaneous” indicates that sensory conditions were not specified and may not have been controlled. Studies were in anesthetized animals except “5” in awake cats, and “7” and “14” in vitro during efforts to mimic natural conditions. Data was obtained with extracellular recordings, except intracellular recordings in 6–7 and 12–14. In “6” and “10,” O/I are only presented for those neurons that displayed “failures” (O/I < 1.0), whereas O/I was reported as 1.0 in the remainder of 24 total neurons in “6,” and in 63, 21, and 9 other cells in “10” in the cases of spontaneous activity in MNTB and AVCN, and sound evoked activity in AVCN, respectively.
Figure 4

Output to input ratios from the references in Table 1. Symbols indicate mean across cells (large circles), individual cells (small circles), s.d. or quantiles (short lines), and range (long lines). For references 6 and 10, the “X” (mean) and other symbols correspond to a subset of cells in which O/I was less than 1, whereas large circles indicate our estimates of the total population mean. The strongest stimulus condition is shown for those references for which more than one condition is reported in Table 1. Reference 1 has data for cat (left) and monkey (right). A single cell with O/I of 150% is not shown for “12”.
Retinal spikes are the cause of all spikes in LGN (Kaplan and Shapley,
Figure 5

Mean SP across 15 neurons as a function of time elapsed since the last EPSP (adapted from Sincich et al.,
Sincich et al. (
Given sufficient knowledge of natural patterns, the theory should allows us to predict membrane excitability and the properties of specific ion channels. In work presented in preliminary form and currently under review for publication, we provide evidence in brain slices of LGN that T-type Ca2+ channels promote predictive homeostasis by restoring SP towards 0.5 during naturally occurring periods of hyperpolarization like that illustrated in Figure 3 (Hong et al.,
Evidence from other types of neurons
In the medial nucleus of the trapezoid body (MNTB) in the early auditory system, the calyx of Held is an exceptionally powerful synapse in which a single presynaptic terminal wraps almost entirely around the postsynaptic soma. It was long believed to be a reliable “fail-safe relay” with SP very near 1. However, the existence of powerful synaptic inhibition suggests that SP should sometimes be lower (Awatramani et al.,
The endbulb of Held in the anteroventral cochlear nucleus (AVCN) is another powerful axosomatic synapse, and SP has been found to be near 0.5 during acoustic stimulation (Englitz et al.,
That even the most powerful axosomatic synapses appear to have average SP nearer to 0.5 than 1.0 under natural conditions supports the theory, and makes it doubtful whether average SP is near 1.0 at any synapse. The climbing fiber to Purkinje cell synapse in the cerebellum is reliable (SP = 1) under standard conditions in brain slices, but the strength of other powerful synapses has been found to be much weaker under natural conditions in vivo, due to both pre- and postsynaptic factors (see Borst,
A typical cerebellar granule cell has four excitatory synapses of approximately equal strength, whereas sympathetic ganglion neurons are autonomic motoneurons that receive synaptic excitation from about 2–5 cholinergic preganglionic neurons, one of which is typically the main driver. Intracellular recordings have found SP to be near 0.5 in each of these cell types (Table 1; Figure 4; Chadderton et al.,
In cells with larger numbers of drivers, such as cortical neurons, unitary EPSPs in the soma are typically not more than 1 mV, and significant spatial summation across synapses is required to evoke a spike. According to the theory, small unitary EPSPs should not be common under natural conditions, since that would cause SP to be very low. Indeed, there can be a high degree of synchronous excitation causing large EPSPs in at least some cortical neurons (DeWeese and Zador,
We have summarized above all of the highly relevant data that we found from simultaneous recordings of synaptic drive and spike output, and it supports our hypothesis that average SP is near 0.5 under natural conditions. However, of the modest number of in vivo studies that are most relevant to our hypothesis, few reported direct and detailed analyses of SP, and few attempted to study natural patterns of synaptic excitation.
Vesicle release probability at presynaptic terminals
When a spike arrives at a single release site in a presynaptic axon terminal, it typically causes release of just 0 or 1 vesicle. Release is commonly said to be “stochastic”, but it is a highly regulated process (Branco and Staras,
Spikes entering a terminal from the axon would presumably be the sole driver and proximal cause of almost all vesicle release. They act via a transient increase in calcium conductance (analogous to an EPSG) and cytosolic concentration (analogous to an EPSP). The free calcium concentration varies from spike to spike and depends on a variety of factors, including the state of potassium channels (that influence action potential duration), calcium channels, and calcium buffers (Meinrenken et al.,
If we could measure RP in real time at a single release site, then we would expect it to have the general features of SP, with numerous processes of excitation and inhibition (“facilitation and depression”) acting on numerous timescales. The goal of predictive homeostasis would be to make accurate predictions of calcium conductance, with the result that vesicle release is maximally sensitive to small variations in the amplitude of calcium events (conductance and concentration). A perfectly accurate expectation would result in RP of 0.5. Like SP, we would expect RP to vary dynamically from nearly 0 to 1 with a homeostatic average of 0.5.
Average RP (across release sites and over time) has been found to fall within an intermediate range (∼0.1–0.9) with an “average of averages” near 0.5 across a diversity of synapses (Branco and Staras,
We would like to know RP as a function of ISI, analogous to SP in Figure 5. However, studies often track the dynamics of RP (through its influence on EPSG amplitude) without quantifying its actual value. Because calcium concentration depends strongly on ISIs, and predictions should be accurate in the case of commonly occurring ISIs, we would expect that RP should be particularly sensitive to variations near these typical intervals. Indeed, in neurons of the fish electrosensory system that have average firing rates near 200 Hz, RP is highly sensitive to ISI variations near the average of 5 ms (Khanbabaie et al.,
A unitary EPSG usually involves multiple release sites, and the powerful synapses discussed above have hundreds. Thus a presynaptic spike will almost always cause an EPSG, and its amplitude will be determined by the average RP across multiple release sites. Because homeostatic mechanisms at each release site are all working independently to “pull” RP towards 0.5, and because each release site experiences the same pattern of spikes, average RP across a presynaptic neuron’s synapses (and the corresponding postsynaptic EPSG amplitude) should not show large variations over time (depression and facilitation should be modest). This has indeed been found to be the case in vivo (Borst,
Conclusion
We believe that taking the neuron’s point of view is essential if we are to understand the relation between biophysical mechanisms and their information content. Perhaps the most striking insight provided by our approach is its simple explanation of why spikes appear so “random”. What appears noisy and even pathological from the standard perspective of a physiologist can be readily understood as optimal from the perspective of a neuron. Taking the neuron’s point of view is a radical departure from past efforts to understand the neural “code”, and it could provide the foundation for a general theory of brain function that is able in principle to explain and predict the diversity of neural mechanisms (Fiorillo,
Statements
Author contributions
Christopher D. Fiorillo conceived the ideas and wrote the manuscript. Jaekyung K. Kim performed the simulations shown in Figures 1 and 2, and Jaekyung K. Kim and Christopher D. Fiorillo designed Figures 1 and 2. Su Z. Hong and Christopher D. Fiorillo reviewed the experimental literature on SP and designed Table 1 and Figure 4.
Acknowledgments
We thank Sunil Kim and Fleur Zeldenrust for helpful comments on the manuscript, and Lawrence Sincich for advice on describing and reproducing his data (Figure 5). Research was supported by a grant to Christopher D. Fiorillo from the Basic Science Research Program (2012R1A1A2006996) of the National Research Foundation of Korea (NRF) funded by the Ministry of Education, Science and Technology.
Conflict of interest
We have cited a patent that discusses the distinction between sensory and motor neurons within its specification (Fiorillo,
References
1
AttwellD.LaughlinS. B. (2001). An energy budget for signaling in the grey matter of the brain. J. Cereb. Blood Flow Metab.21, 1133–1145. 10.1097/00004647-200110000-00001
2
AvissarM.FurmanA. C.SaundersJ. C.ParsonsT. D. (2007). Adaptation reduces spike-count reliability, but not spike-timing precision, of auditory nerve responses. J. Neurosci.27, 6461–6472. 10.1523/jneurosci.5239-06.2007
3
AwatramaniG. B.TurecekR.TrussellL. O. (2004). Inhibitory control at a synaptic relay. J. Neurosci.24, 2643–2647. 10.1523/jneurosci.5144-03.2004
4
BairW.KochC. (1996). Temporal precision of spike trains in extrastriate cortex of the behaving macaque monkey. Neural Comput.8, 1185–1202. 10.1162/neco.1996.8.6.1185
5
BarlowH. B. (1961). “Possible principles underlying the transformation of sensory messages,” in Sensory Communication, ed RosenblithW. A. (Cambridge, MA: MIT Press), 217–234.
6
BeckJ. M.MaW. J.KianiR.HanksT.ChurchlandA. K.RoitmanJ.et al. (2008). Probabilistic population codes for Bayesian decision making. Neuron60, 1142–1152. 10.1016/j.neuron.2008.09.021
7
BergR. W.AlaburdaA.HounsgaardJ. (2007). Balanced inhibition and excitation drive spike activity in spinal half-centers. Science315, 390–393. 10.1126/science.1134960
8
BerkesP.OrbánG.LengyelM.FiserJ. (2011). Spontaneous cortical activity reveals hallmarks of an optimal internal model of the environment. Science331, 83–87. 10.1126/science.1195870
9
BerryM. J.WarlandD. K.MeisterM. (1997). The structure and precision of retinal spike trains. Proc. Natl. Acad. Sci. U S A94, 5411–5416. 10.1073/pnas.94.10.5411
10
BienenstockE. L.CooperL. N.MunroP. W. (1982). Theory for the development of neuron selectivity: orientation specificity and binocular interaction in visual cortex. J. Neurosci.2, 32–48.
11
BlitzD. M.RegehrW. G. (2005). Timing and specificity of feed-forward inhibition within the LGN. Neuron45, 917–928. 10.1016/j.neuron.2005.01.033
12
BorstJ. G. G. (2010). The low synaptic release probability in vivo. Trends Neurosci.33, 259–266. 10.1016/j.tins.2010.03.003
13
BrancoT.HäusserM. (2011). Synaptic integration gradients in single cortical pyramidal cell dendrites. Neuron69, 885–892. 10.1016/j.neuron.2011.02.006
14
BrancoT.StarasK. (2009). The probability of neurotransmitter release: variability and feedback control at single synapses. Nat. Rev. Neurosci.10, 373–383. 10.1038/nrn2634
15
BrattonB.DaviesP.JänigW.McAllenR. (2010). Ganglionic transmission in a vasomotor pathway studied in vivo. J. Physiol.588, 1647–1659. 10.1113/jphysiol.2009.185025
16
BrennerN.BialekW.de Ruyter van SteveninckR. (2000). Adaptive rescaling maximizes information transmission. Neuron26, 695–702. 10.1016/s0896-6273(00)81205-2
17
ButtsD. A.WengC.JinJ.YehC.-I.LesicaN. A.AlonsoJ.-M.et al. (2007). Temporal precision in the neural code and the timescales of natural vision. Nature449, 92–95. 10.1038/nature06105
18
CalvinW. H.StevensC. F. (1968). Synaptic noise and other sources of randomness in motoneuron interspike intervals. J. Neurophysiol.31, 574–587.
19
CarandiniM. (2004). Amplification of trial-to-trial response variability by neurons in visual cortex. PLoS Biol.2:e264. 10.1371/journal.pbio.0020264
20
CarandiniM.HortonJ. C.SincichL. C. (2007). Thalamic filtering of retinal spike trains by postsynaptic summation. J. Vis.7, 20–20. 10.1167/7.14.20
21
CastiA.HayotF.XiaoY.KaplanE. (2008). A simple model of retina-LGN transmission. J. Comput. Neurosci.24, 235–252. 10.1007/s10827-007-0053-7
22
ChaddertonP.MargrieT. W.HäusserM. (2004). Integration of quanta in cerebellar granule cells during sensory processing. Nature428, 856–860. 10.1038/nature02442
23
ChurchlandM. M.YuB. M.CunninghamJ. P.SugrueL. P.CohenM. R.CorradoG. S.et al. (2010). Stimulus onset quenches neural variability: a widespread cortical phenomenon. Nat. Neurosci.13, 369–378. 10.1038/nn.2501
24
ClarkA. (2013). Whatever next? Predictive brains, situated agents and the future of cognitive science. Behav. Brain Sci.36, 181–204. 10.1017/s0140525x12000477
25
CooperL. N.BearM. F. (2012). The BCM theory of synapse modification at 30: interaction of theory with experiment. Nat. Rev. Neurosci.13, 798–810. 10.1038/nrn3353
26
DanY.AtickJ. J.ReidR. C. (1996). Efficient coding of natural scenes in the lateral geniculate nucleus: experimental test of a computational theory. J. Neurosci.16, 3351–3362.
27
DeneveS. (2008a). Bayesian spiking neurons I: inference. Neural Comput.20, 91–117. 10.1162/neco.2008.20.1.91
28
DeneveS. (2008b). Bayesian spiking neurons II: learning. Neural Comput.20, 118–145. 10.1162/neco.2008.20.1.118
29
DeneveS.LathamP. E.PougetA. (2001). Efficient computation and cue integration with noisy population codes. Nat. Neurosci.4, 826–831. 10.1038/90541
30
de Ruyter van SteveninckR. R.LewenG. D.StrongS. P.KoberleR.BialekW. (1997). Reproducibility and variability in neural spike trains. Science275, 1805–1808. 10.1126/science.275.5307.1805
31
DeWeeseM. R.WehrM.ZadorA. M. (2003). Binary spiking in auditory cortex. J. Neurosci.23, 7940–7949.
32
DeWeeseM. R.ZadorA. M. (2006). Non-Gaussian membrane potential dynamics imply sparse, synchronous activity in auditory cortex. J. Neurosci.26, 12206–12218. 10.1523/jneurosci.2813-06.2006
33
EnglitzB.TolnaiS.TypltM.JostJ.RübsamenR. (2009). Reliability of synaptic transmission at the synapses of held in vivo under acoustic stimulation. PLoS One4:e7014. 10.1371/journal.pone.0007014
34
FiorilloC. D. (2008). Towards a general theory of neural computation based on prediction by single neurons. PLoS One3:e3298. 10.1371/journal.pone.0003298
35
FiorilloC. D. (2012). Beyond Bayes: on the need for a unified and Jaynesian definition of probability and information within neuroscience. Information3, 175–203. 10.3390/info3020175
36
FiorilloC. D. (2013a). Prediction by single neurons. United States Patent 8504502.
37
FiorilloC. D. (2013b). Two dimensions of value: dopamine neurons represent reward but not aversiveness. Science341, 546–549. 10.1126/science.1238699
38
FiorilloC. D.SongM. R.YunS. R. (2013). Multiphasic temporal dynamics in responses of midbrain dopamine neurons to appetitive and aversive stimuli. J. Neurosci.33, 4710–4725. 10.1523/jneurosci.3883-12.2013
39
FrémauxN.SprekelerH.GerstnerW. (2010). Functional requirements for reward-modulated spike-timing-dependent plasticity. J. Neurosci.30, 13326–13337. 10.1523/jneurosci.6249-09.2010
40
FristonK. J. (2010). The free-energy principle: a unified brain theory?Nat. Rev. Neurosci.11, 127–138. 10.1038/nrn2787
41
FristonK. J.ThorntonC.ClarkA. (2012). Free-energy minimization and the dark-room problem. Front. Psychol.3:130. 10.3389/fpsyg.2012.00130
42
GentetL. J.AvermannM.MatyasF.StaigerJ. F.PetersenC. C. H. (2010). Membrane potential dynamics of GABAergic neurons in the barrel cortex of behaving mice. Neuron65, 422–435. 10.1016/j.neuron.2010.01.006
43
HermannJ.PeckaM.von GersdorffH.GrotheB.KlugA. (2007). Synaptic transmission at the calyx of held under in vivo-like activity levels. J. Neurophysiol.98, 807–820. 10.1152/jn.00355.2007
44
HigleyM. J.ContrerasD. (2006). Balanced excitation and inhibition determine spike timing during frequency adaptation. J. Neurosci.26, 448–457. 10.1523/jneurosci.3506-05.2006
45
HongS.KimH. R.YunS. R.FiorilloC. D. (2013). “T-type calcium channels promote predictive homeostasis in thalamocortical neurons of LGN,” in Computational and Systems Neuroscience 2013 (COSYNE) (Salt Lake City, UT, USA).
46
HosoyaT.BaccusS. A.MeisterM. (2005). Dynamic predictive coding by the retina. Nature436, 71–77. 10.1038/nature03689
47
HoyerP. O.HyvarinenA. (2003). “Interpreting neural response variability as Monte Carlo sampling of the posterior,” in Advances in Neural Information Processing Systems 15, eds BeckerS. et al. (Cambridge, MA: MIT Press), 277–284.
48
IzhikevichE. M. (2006). Solving the distal reward problem through linkage of STDP and dopamine signaling. Cereb. Cortex17, 2443–2452. 10.1093/cercor/bhl152
49
JaynesE. T. (2003). Probability Theory: The Logic of Science.Cambridge, UK: Cambridge University Press.
50
JazayeriM.MovshonJ. A. (2006). Optimal representation of sensory information by neural populations. Nat. Neurosci.9, 690–696. 10.1038/nn1691
51
KaplanE.PurpuraK.ShapleyR. M. (1987). Contrast affects the transmission of visual information through the mammalian lateral geniculate nucleus. J. Physiol.391, 267–288.
52
KaplanE.ShapleyR. (1984). The origin of the S (slow) potential in the mammalian lateral geniculate nucleus. Exp. Brain Res.55, 111–116. 10.1007/BF00240504
53
KhanbabaieR.NesseW. H.LongtinA.MalerL. (2010). Kinetics of fast short-term depression are matched to spike train statistics to reduce noise. J. Neurophysiol.103, 3337–3348. 10.1152/jn.00117.2010
54
Kopp-ScheinpflugC.LippeW. R.DorrscheidtG. J.RübsamenR. (2003). The medial nucleus of the trapezoid body in the gerbil is more than a relay: comparison of pre- and postsynaptic activity. J. Assoc. Res. Otolaryngol.4, 1–23. 10.1007/s10162-002-2010-5
55
KördingK. P.WolpertD. M. (2004). Bayesian integration in sensorimotor learning. Nature427, 244–247. 10.1038/nature02169
56
KuenzelT.BorstJ. G. G.van der HeijdenM. (2011). Factors controlling the input-output relationship of spherical bushy cells in the gerbil cochlear nucleus. J. Neurosci.31, 4260–4273. 10.1523/jneurosci.5433-10.2011
57
LaughlinS. B. (1981). A simple coding procedure enhances a neuron’s information capacity. Z. Naturforsch. C36, 910–912.
58
LinskerR. (1988). Self-organization in a perceptual network. Computer21, 105–117. 10.1109/2.36
59
LondonM.RothA.BeerenL.HäusserM.LathamP. E. (2010). Sensitivity to perturbations in vivo implies high noise and suggests rate coding in cortex. Nature466, 123–127. 10.1038/nature09086
60
LorteijeJ. A. M.RusuS. I.KushmerickC.BorstJ. G. G. (2009). Reliability and precision of the mouse calyx of Held synapse. J. Neurosci.29, 13770–13784. 10.1523/jneurosci.3285-09.2009
61
MaW. J.BeckJ. M.LathamP. E.PougetA. (2006). Bayesian inference with probabilistic population codes. Nat. Neurosci.9, 1432–1438. 10.1038/nn1790
62
MainenZ. F.SejnowskiT. J. (1995). Reliability of spike timing in neocortical neurons. Science268, 1503–1506. 10.1126/science.7770778
63
MarderE.BucherD. (2007). Understanding circuit dynamics using the stomatogastric nervous system of lobsters and crabs. Annu. Rev. Physiol.69, 291–316. 10.1146/annurev.physiol.69.031905.161516
64
MastronardeD. N. (1989). Correlated firing of retinal ganglion cells. Trends Neurosci.12, 75–80. 10.1016/0166-2236(89)90140-9
65
McAllenR. M.SaloL. M.PatonJ. F. R.PickeringA. E. (2011). Processing of central and reflex vagal drives by rat cardiac ganglion neurones: an intracellular analysis. J. Physiol.589, 5801–5818. 10.1113/jphysiol.2011.214320
66
McLaughlinM.van der HeijdenM.JorisP. X. (2008). How secure is in vivo synaptic transmission at the calyx of Held?J. Neurosci.28, 10206–10219. 10.1523/JNEUROSCI.2735-08.2008
67
MeinrenkenC. J.BorstJ. G. G.SakmannB. (2003). Local routes revisited: the space and time dependence of the Ca2+ signal for phasic transmitter release at the rat calyx of Held. J. Physiol.547, 665–689. 10.1113/jphysiol.2002.032714
68
MiglioreM.HoffmanD. A.MageeJ. C.JohnstonD. (1999). Role of an A-type K+ conductance in the back-propagation of action potentials in the dendrites of hippocampal pyramidal neurons. J. Comput. Neurosci.7, 5–15. 10.1023/A:1008906225285
69
MovshonJ. A.KiorpesL.HawkenM. J.CavanaughJ. R. (2005). Functional maturation of the macaque’s lateral geniculate nucleus. J. Neurosci.25, 2712–2722. 10.1523/jneurosci.2356-04.2005
70
NiemeierM.CrawfordJ. D.TweedD. B. (2003). Optimal transsaccadic integration explains distorted spatial perception. Nature422, 76–80. 10.1038/nature01439
71
OkunM.LamplI. (2008). Instantaneous correlation of excitation and inhibition during ongoing and sensory-evoked activities. Nat. Neurosci.11, 535–537. 10.1038/nn.2105
72
PougetA.DayanP.ZemelR. (2000). Information processing with population codes. Nat. Rev. Neurosci.1, 125–132. 10.1038/35039062
73
PouilleF.ScanzianiM. (2001). Enforcement of temporal fidelity in pyramidal cells by somatic feed-forward inhibition. Science293, 1159–1163. 10.1126/science.1060342
74
PouletJ. F. A.PetersenC. C. H. (2008). Internal brain state regulates membrane potential synchrony in barrel cortex of behaving mice. Nature454, 881–885. 10.1038/nature07150
75
RiekeF.WarlandD.de Ruyter van SteveninckR. R.BialekW. (1997). Spikes: Exploring the Neural Code.Cambridge, MA: MIT Press.
76
SchultzW.DayanP.MontagueP. R. (1997). A neural substrate of prediction and reward. Science275, 1593–1599. 10.1126/science.275.5306.1593
77
SenguptaB.LaughlinS. B.NivenJ. E. (2013). Balanced excitatory and inhibitory synaptic currents promote efficient coding and metabolic efficiency. PLoS Comput. Biol.9:e1003263. 10.1371/journal.pcbi.1003263
78
ShadlenM. N.NewsomeW. T. (1994). Noise, neural codes and cortical organization. Curr. Opin. Neurobiol.4, 569–579. 10.1016/0959-4388(94)90059-0
79
ShadlenM. N.NewsomeW. T. (1998). The variable discharge of cortical neurons: implications for connectivity, computation and information coding. J. Neurosci.18, 3870–3896.
80
ShannonC. E. (1948). A mathematical theory of communication. Bell Syst. Tech. J.27, 379–423. 10.1002/j.1538-7305.1948.tb01338.x
81
SharpeeT. O.SugiharaH.KurganskyA. V.RebrikS. P.StrykerM. P.MillerK. D. (2006). Adaptive filtering enhances information transmission in visual cortex. Nature439, 936–942. 10.1038/nature04519
82
ShermanS. M.GuilleryR. W. (1998). On the actions that one nerve cell can have on another: distinguishing “drivers” from ‘modulators’. Proc. Natl. Acad. Sci. U S A95, 7121–7126. 10.1073/pnas.95.12.7121
83
ShuY.HasenstaubA.McCormickD. A. (2003). Turning on and off recurrent balanced cortical activity. Nature423, 288–293. 10.1038/nature01616
84
SimoncelliE. P.OlshausenB. A. (2001). Natural image statistics and neural representation. Annu. Rev. Neurosci.24, 1193–1216. 10.1146/annurev.neuro.24.1.1193
85
SincichL. C.AdamsD. L.EconomidesJ. R.HortonJ. C. (2007). Transmission of spike trains at the retinogeniculate synapse. J. Neurosci.27, 2683–2692. 10.1523/jneurosci.5077-06.2007
86
SoftkyW. R. (1995). Simple codes versus efficient codes. Curr. Opin. Neurobiol.5, 239–247. 10.1016/0959-4388(95)80032-8
87
SoftkyW. R.KochC. (1993). The highly irregular firing of cortical cells is inconsistent with temporal integration of random EPSPs. J. Neurosci.13, 334–350.
88
SrinivasanM. V.LaughlinS. B.DubsA. (1982). Predictive coding: a fresh view of inhibition in the retina. Proc. R. Soc. Lond. B Biol. Sci.216, 427–459. 10.1098/rspb.1982.0085
89
StemmlerM.KochC. (1999). How voltage-dependent conductances can adapt to maximize the information encoded by neuronal firing rate. Nat. Neurosci.2, 521–527. 10.1038/9173
90
TolhurstD. J.MovshonJ. A.DeanA. F. (1983). The statistical reliability of signals in single neurons in cat and monkey visual cortex. Vision Res.23, 775–785. 10.1016/0042-6989(83)90200-6
91
UsreyW. M.ReppasJ. B.ReidR. C. (1998). Paired-spike interactions and synaptic efficacy of retinal inputs to the thalamus. Nature395, 384–387. 10.1038/26487
92
WangX.HirschJ. A.SommerF. T. (2010b). Recoding of sensory information across the retinothalamic synapse. J. Neurosci.30, 13567–13577. 10.1523/JNEUROSCI.0910-10.2010
93
WangH. P.SpencerD.FellousJ. M.SejnowskiT. J. (2010a). Synchrony of thalamocortical inputs maximizes cortical reliability. Science328, 106–109. 10.1126/science.1183108
94
WangX.SommerF. T.HirschJ. A. (2011). Inhibitory circuits for visual processing in thalamus. Curr. Opin. Neurobiol.21, 726–733. 10.1016/j.conb.2011.06.004
95
WangX.-J.LiuY.Sanchez-VivesM. V.McCormickD. A. (2003). Adaptation and temporal decorrelation by single neurons in the primary visual cortex. J. Neurophysiol.89, 3279–3293. 10.1152/jn.00242.2003
96
WangX.WeiY.VaingankarV.WangQ.KoepsellK.SommerF. T.et al. (2007). Feedforward excitation and inhibition evoke dual modes of firing in the cat’s visual thalamus during naturalistic viewing. Neuron55, 465–478. 10.1016/j.neuron.2007.06.039
97
WehrM.ZadorA. M. (2003). Balanced inhibition underlies tuning and sharpens spike timing in auditory cortex. Nature426, 442–446. 10.1038/nature02116
98
WeissY.SimoncelliE. P.AdelsonE. H. (2002). Motion illusions as optimal percepts. Nat. Neurosci.5, 598–604. 10.1038/nn858
99
WeyandT. G. (2007). Retinogeniculate transmission in wakefulness. J. Neurophysiol.98, 769–785. 10.1152/jn.00929.2006
100
WilentW. B.ContrerasD. (2005). Dynamics of excitation and inhibition underlying stimulus selectivity in rat somatosensory cortex. Nat. Neurosci.8, 1364–1370. 10.1038/nn1545
101
YangZ.PurvesD. (2003). A statistical explanation of visual space. Nat. Neurosci.6, 632–640. 10.1038/nn1059
102
YuJ.FersterD. (2010). Membrane potential synchrony in primary visual cortex during sensory stimulation. Neuron68, 1187–1201. 10.1016/j.neuron.2010.11.027
Summary
Keywords
Bayesian, Jaynes, inference, prediction, neural code, noise, stochastic, random
Citation
Fiorillo CD, Kim JK and Hong SZ (2014) The meaning of spikes from the neuron’s point of view: predictive homeostasis generates the appearance of randomness. Front. Comput. Neurosci. 8:49. doi: 10.3389/fncom.2014.00049
Received
04 November 2013
Accepted
02 April 2014
Published
29 April 2014
Volume
8 - 2014
Edited by
Rava Azeredo Da Silveira, Ecole Normale Supérieure, France
Reviewed by
Benjamin Torben-Nielsen, The Hebrew University of Jerusalem, Israel; Gianluigi Mongillo, Paris Descartes University, France
Copyright
© 2014 Fiorillo, Kim and Hong.
This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.
Correspondence: Christopher D. Fiorillo, Department of Bio and Brain Engineering, Korean Advanced Institute of Science and Engineering (KAIST), 291 Daehak-ro, Building 611, Yuseong-gu, Daejeon 305-701, South Korea e-mail: fiorillo@kaist.ac.kr
This article was submitted to the journal Frontiers in Computational Neuroscience.
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.