Abstract
In the prefrontal cortex (PFC), higher-order cognitive functions and adaptive flexible behaviors rely on continuous dynamical sequences of spiking activity that constitute neural trajectories in the state space of activity. Neural trajectories subserve diverse representations, from explicit mappings in physical spaces to generalized mappings in the task space, and up to complex abstract transformations such as working memory, decision-making and behavioral planning. Computational models have separately assessed learning and replay of neural trajectories, often using unrealistic learning rules or decoupling simulations for learning from replay. Hence, the question remains open of how neural trajectories are learned, memorized and replayed online, with permanently acting biological plasticity rules. The asynchronous irregular regime characterizing cortical dynamics in awake conditions exerts a major source of disorder that may jeopardize plasticity and replay of locally ordered activity. Here, we show that a recurrent model of local PFC circuitry endowed with realistic synaptic spike timing-dependent plasticity and scaling processes can learn, memorize and replay large-size neural trajectories online under asynchronous irregular dynamics, at regular or fast (sped-up) timescale. Presented trajectories are quickly learned (within seconds) as synaptic engrams in the network, and the model is able to chunk overlapping trajectories presented separately. These trajectory engrams last long-term (dozen hours) and trajectory replays can be triggered over an hour. In turn, we show the conditions under which trajectory engrams and replays preserve asynchronous irregular dynamics in the network. Functionally, spiking activity during trajectory replays at regular timescale accounts for both dynamical coding with temporal tuning in individual neurons, persistent activity at the population level, and large levels of variability consistent with observed cognitive-related PFC dynamics. Together, these results offer a consistent theoretical framework accounting for how neural trajectories can be learned, memorized and replayed in PFC networks circuits to subserve flexible dynamic representations and adaptive behaviors.
Introduction
As when a few introductory notes recall a melody, in the immense space of known melodies, cerebral networks are able to memorize and replay complex temporal patterns in a flexible way. Such temporal patterns rely on continuous dynamical sequences of spiking activity, i.e., neural trajectories, that occur in recurrent neural networks of the prefrontal cortex (PFC) (Bakhurin et al., ; Paton and Buonomano, 2018; Wang et al., 2018). These neural trajectories emerge with learning, relying on dynamical engrams, which distinguish them from classical static engrams underlying Hebbian neuronal assemblies. In turn, these engrams likely arise through activity-dependent synaptic plasticity (Goto et al., ; Bittner et al., ). Hence, a robust understanding of the interplay between prefrontal dynamics and biological plastic processes is necessary to understand the emergence of functional neural trajectories and engrams. In the PFC of behaving animals, neural trajectories are embedded in an asynchronous and irregular background state activity that is markedly disordered (Destexhe et al., ; London et al., ). However, how synaptic plasticity builds engrams that are not erased by spontaneous activity and yet are not strong enough to alter irregular PFC dynamics remains an open question.
Neural trajectories correspond to organized spatio-temporal representations that peregrinate within the neural space (Shenoy et al., 2013). They are prominent in prefrontal cortices (Mante et al., 2013), where they subserve higher-order cognitive functions at diverse levels of abstraction (Wutz et al., 2018). In prefrontal areas, at the lowest levels of abstraction, neural trajectories can map the actual animal's position during effective trajectories within explicit spaces during visual perception (Mante et al., 2013) or navigation (Fujisawa et al., ; Zielinski et al., 2019). Beyond spatial mapping, neural trajectories can also depict generalized topological locations that are isomorphic to the task space, by multiplexing position, representation of goal locations and choice-related information (Fujisawa et al., ; Mashhoori et al., 2018; Yu et al., 2018; Kaefer et al., ). Neural trajectories have also been shown to subserve dynamical coding and manipulation of information during delay activities in working memory tasks involving the PFC (Lundqvist et al., ). In this context, neural trajectories do not represent explicit trajectories in external spaces, but implicit representations—of ongoing information and cognitive operations—that may prove useful for the task.
Rather than static maintenance of persistent activity in a group of cells, many working-memory representations unfold in the space of neural activity under the form of continuous trajectories, as neurons successively activate in “relay races” sequences of transient activity (Batuev, ; Brody et al., ; Cromer et al., ; Yang et al., 2014; Schmitt et al., 2017; Enel et al., ). In the PFC, neural trajectories can form the substrate for dynamic (Sreenivasan et al., 2014) but also, counterintuitively, for stable representations (Druckmann and Chklovskii, ). Neural trajectory-mediated dynamical representations can subserve the retrospective working memory of spatial (Batuev, ; Yang et al., 2014) or quantitative (Brody et al., ) cues, symbolic categories (Cromer et al., ), values (Enel et al., ), or behavioral rules (Schmitt et al., 2017). They can also serve prospective working memory in computational processes transforming previously encoded information, such as, for e.g., in visuo-motor transformations (Spaak et al., 2017), in the representation of elapsed time (Tiganj et al., 2017) or in the encoding of forthcoming behaviors (Fujisawa et al., ; Ito et al., ; Nakajima et al., 2019; Passecker et al., 2019). Neural trajectories in the neural space can also appear as sequences of states that involve combinations of active neurons (Batuev, ; Abeles et al., ; Seidemann et al., 1996; La Camera et al., ). Thus, neural trajectories appear in diverse forms and in different functional contexts where they can map actual trajectories in external spaces, remember previously encountered trajectories, or predict forthcoming trajectories during active computational processes requiring dynamical representations.
Neural trajectories in the PFC are adaptive (Euston et al., ; Mante et al., 2013): they are learned and memorized, to be “replayed” later. The timescale of the replay depends on the behavioral context. Regular timescale replays operate at the behavioral timescale, lasting seconds (Batuev, ; Fujisawa et al., ; Cromer et al., ; Mante et al., 2013; Yang et al., 2014; Ito et al., ; Schmitt et al., 2017; Tiganj et al., 2017; Nakajima et al., 2019; Passecker et al., 2019; Enel et al., ). Thus, such replays unfold online as current behavior is executed in interaction with the external world, to subserve retrospective working memory of past information, on-going dynamical computations, or prospective representation of forthcoming behaviors. Typically, regular replays are triggered by behaviorally–relevant external events (e.g., cues or go signals in working memory tasks, or the current position in navigational tasks). Some replays that may appear as spontaneous can be presumably triggered by internal self-paced decision signals within the PFC (e.g., choices). In all cases, such triggered regular replays rely on internal mechanisms within PFC circuits allowing for the autonomous propagation of proper sequences of activity, once initial neurons of the neural trajectory have been triggered. A major goal of the present study is to decipher how plastic processes allow PFC circuits to learn and replay trajectories, i.e., autonomously generate neural trajectory completion, based on an initial trigger.
Besides, fast timescale replays exist that last a few hundred milliseconds during awake (Jadhav et al., ; Mashhoori et al., 2018; Yu et al., 2018; Shin et al., 2019; Kaefer et al., ) and sleeping (Euston et al., ; Peyrache et al., 2009) states. Beyond their much shorter duration, PFC fast replays are distinct from regular ones, in that they typically operate offline and often co-occur with fast replays in the hippocampal CA1 field (Jadhav et al., ). Replay activity in PFC and CA1 presents high degrees of task-dependent spatial and temporal correlations (Jadhav et al., ; Yu et al., 2018; Shin et al., 2019), subserving functional coordination combining metric (hippocampus) and task-related (PFC) spatial representations (Pfeiffer and Foster, 2013; Zielinski et al., 2019). These fast replays occur during sharp-wave ripples (SWR) episodes (Jadhav et al., ; Yu et al., 2018; Shin et al., 2019), which represent critical events for behavioral learning (Jadhav et al., ) and during which animals forge forthcoming decisions (choices, trajectories, for e.g., Jadhav et al., ; Mashhoori et al., 2018; Kaefer et al., ), based on the recall of past experiences (actions, trajectories, outcomes, for e.g., Jadhav et al., ; Mashhoori et al., 2018). Such coordination across both structures presumably emerges through their reciprocal, direct and indirect, synaptic interactions (Witter and Amaral, 2004). Different studies have pointed out information flow biases from CA1 to PFC (Jadhav et al., ) or from PFC to CA1 (Ito et al., ) directions, depending on behavioral contexts. However, SWR-related replays in the hippocampus correlate with fast replays in reduced subsets of PFC neurons (Jadhav et al., ; Yu et al., 2018) that carry generalized spatial representations but not specific trajectories (Yu et al., 2018). Moreover, fast timescale PFC replays are independent of hippocampal replays during computational processes inherent to the PFC, such as rule switching tasks (Kaefer et al., ). Therefore, as for regular replays, we examined how plastic processes allow for the emergence of fast timescale replays autonomously within local recurrent PFC circuits.
Neuronal trajectories consist of robust forms of ordered local activity occurring within a disordered global activity, i.e., the chaotic, asynchronous irregular (AI) state characteristic of the prefrontal cortex in the waking state (Destexhe et al., ; London et al., ). This coexistence poses a problem at the plasticity level, because the noisy AI regime constitutes a potential source of perturbation for synaptic engrams (Boustani et al., ; Litwin-Kumar and Doiron, ), whereas strengthened connectivity pathways may exert a synchronizing influence on the network, dramatically altering the chaotic nature of background activity. However, there is currently no biophysically-grounded theoretical framework accounting for the way neural trajectories are learned, memorized and replayed within recurrent cortical networks. In principle, synaptic plasticity, a major substrate of learning, may sculpt oriented connective pathways promoting the propagation of neuronal trajectories, because modifications of synaptic connections are activity-dependent. Specifically, the sequential activation of differentially tuned neurons during successively crossed spatial positions (during navigational trajectories) or representational states (during dynamical cognitive processes) could strengthen connections between neurons, creating oriented pathways (referred to as trajectory engrams hereafter) within recurrent cortical networks. If sufficiently strengthened, engrams could allow the propagation of packets of neuronal activity along them. From an initial stimulation of neurons located at the beginning of the engram, due to the strong connections linking them in the direction of the trajectory, neurons could reactivate sequentially, i.e., perform trajectory replay.
Recurrent neural network models have shown that activity-dependent synaptic plasticity rules can enable the formation of trajectory engrams due to long-term potentiation (LTP) and depression (LTD) together with homeostatic scaling (Liu and Buonomano, ; Clopath et al., ; Fiete et al., ; Klampfl and Maass, ). Moreover, trajectory engrams can propagate neuronal trajectories through sequential activation of neurons in recurrent model networks (Liu and Buonomano, ; Fiete et al., ; Klampfl and Maass, ; Laje and Buonomano, ; Chenkov et al., ). However, the above models of neural trajectories do not elucidate the biological basis of learning and replay in neurophysiological situations encountered by PFC networks for several reasons. First, in these models, trajectory learning is either ignored (hard-written trajectory engram; Chenkov et al., ), unrelated to behavior (random formation of arbitrary trajectory; Liu and Buonomano, ; Fiete et al., ), based on artificial learning rules (Laje and Buonomano, ) or on biophysically unrealistic rules in terms of neuronal activity and synaptic plasticity constraints (Liu and Buonomano, ; Fiete et al., ; Klampfl and Maass, ). Moreover, trajectory replay is absent (Clopath et al., ) or unable to operate from an initial trigger (Klampfl and Maass, ), or the ability to memorize and replay trajectory engrams and replays long-term is not tested (Liu and Buonomano, ; Clopath et al., ; Fiete et al., ; Klampfl and Maass, ; Laje and Buonomano, ; Chenkov et al., ). Finally, none of these models evaluate the capacity for trajectory learning and replay in the realistic context where network activity undergoes AI dynamics, whereas it is characteristic of the awake state in the cortex (Destexhe et al., ; London et al., ). The interactions between synaptic plasticity and AI dynamics has so far only been assessed for static Hebbian engrams (Morrison et al., 2007; Boustani et al., ; Litwin-Kumar and Doiron, ) but not for dynamic trajectories.
The disordered activity of AI cortical dynamics represents a potentially important source of disturbance at many stages. Indeed, AI regime activity may spontaneously engage plastic processes (before any trajectory presentation), affecting the synaptic network matrix, and leading to altered network dynamics with divergence toward silence or saturation (Siri et al., 2007). Noisy activity may also interfere with the learning of the trajectory engram, by adding erratic entries of calcium to trajectory presentation-induced calcium, leading to jeopardized downstream decoding of calcium as well as erratic switches between long-term potentiation (LTP) and long-term depression (LTD) of synaptic weights. After learning, the continuous effects of AI regime activity-induced plastic processes (LTD or scaling) might erase the trajectory engram during memorization and jeopardize trajectory replay through the destabilizing influence of activity noise. On the other side of the interaction, trajectory learning through Hebbian synaptic plasticity may potentially, in turn, seriously disrupt AI regime activity (Morrison et al., 2007; Siri et al., 2007). Therefore, it remains uncertain whether realistic biological synaptic plasticity rules are well-suited for proper learning and memorizing of trajectory engrams as well as replay of learned trajectories in PFC physiological conditions.
Here, we assessed how learning, memorization and replay of trajectories can arise from biologically realistic synaptic learning rules in physiological PFC networks displaying disordered AI regime activity. To do so, we built a local recurrent biophysical network model designed to capture replay events like those observed in the PFC. Although designed to fit PFC collective spontaneous and triggered neural dynamics, its intrinsic, synaptic and architectural properties are shared across other cortices, allowing for generalization of the results to other non-PFC cortical areas displaying replays. The model displayed AI dynamics and was endowed with realistic Hebbian (Hebb, ) spike timing-dependent plasticity (STDP) of excitatory synapses (Bi and Poo, ). Synaptic modifications operate through calcium-signaling dynamics capturing NMDA-dependent non-linear pre- to post-synaptic associativity (Graupner and Brunel, ) and calcium-dependent phosphorylation of synaptic weights with realistic activity-dependent kinase/phosphatase (aKP) dynamics, conferring a rapid, graded and bidirectional induction together with slow maintenance, consistent with learning and memory timescales observed in animal and human (Delord et al., ). Moreover, the model incorporates synaptic scaling, which ensures normalization of pre-synaptic weights, as found in the cortex (Turrigiano et al., 1998; Wang and Gao, 2012; Sweatt, 2016). We show, that, in this realistic model, presenting a stimulus trajectory allowed for rapid learning of a trajectory engram as well as long-term memorization of the trajectory engram despite the disturbing influence of the AI regime. In turn, the STDP learning rule and trajectory engram did not affect the spontaneous AI regime despite their influence on all excitatory neurons from the network. Moreover, we show that trajectory replay accounted for essential aspects of information coding in the PFC, including robustness of replays at the timescale of seconds, fast and regular replays, chunking, large inter-trial variability, and the ability to account for the dual dynamical and persistent aspects of working memory representations.
Materials and Methods
Model of Biophysical Local Recurrent Neural Network
We built a biophysical model of a prefrontal local recurrent neural network, endowed with detailed biological properties of its neurons and connections. While the model is presented as PFC, its synaptic and neural properties are generally preserved across cortical areas, allowing for generalization of the results to non-PFC cortical areas. The network model contained N neurons that were either excitatory (E) or inhibitory (I) (neurons projecting only glutamate or GABA, respectively; Dale, ), with probabilities pE and pI = 1−pE, respectively, and (Beaulieu et al., ). Connectivity was sparse (i.e., only a fraction of all possible connections exists, see pE→E, pE→I, pI→E, pI→I parameter values; Thomson, 2002) with no autapses (self-connections) and EE connections (from E to E neurons) drawn to insure the over-representation of bidirectional connections in cortical networks (four times more than randomly drawn according to a Bernoulli scheme; Song et al., 2005; Wang et al., 2006). The synaptic weights w(i,j) of existing connections were drawn identically and independently from a log-normal distribution of parameters μw and σw (Song et al., 2005).
To cope with simulation times required for the massive explorations ran in the parameter space, neurons were modeled as leaky integrate-and-fire (LIF) neurons. The membrane potential of neuron j followed
where neurons spike when the membrane potential reaches the threshold θ, and repolarization to Vrest occurred after a refractory period ΔtAP.
The leak current followed
where is the maximal conductance and VL the equilibrium potential of the leak current.
The recurrent synaptic current on post-synaptic neuron j, from—either excitatory or inhibitory—pre-synaptic neurons (indexed by i), was
The delay for synaptic conduction and transmission, Δtsyn, was considered uniform across the network (Brunel and Wang, ). Synaptic recurrent currents followed
where w(i,j) is the synaptic weight, px(i) the opening probability of channel-receptors and Vx the reversal potential of the current. The NMDA current followed
incorporating the magnesium block voltage-dependence modeled (Jahr and Stevens, ) as
The channel rise times were approximated as instantaneous (Brunel and Wang, ) and bounded, with first-order decay
where δ is the dirac function and t(i) the times of the pre-synaptic action potentials (APs).
Recurrent excitatory and inhibitory currents were balanced in each post-synaptic neuron (Shu et al., 2003; Haider et al., ; Xue et al., 2014), according to driving forces and excitation/inhibition weight ratio, through
with being an approximation of the average membrane potential.
Furthermore, all recurrent maximal conductances were multiplied by gRec, and by gE→E, gE→I, gI→E or gI→I according to the excitatory or inhibitory nature of pre- and post-synaptic populations.
The feed-forward synaptic current ISyn.FF(j) (putatively arising from sub-cortical and cortical inputs) consisted of an AMPA component.
with a constant opening probability pAMPA.FF.
Synaptic Spike Timing-Dependent Plasticity (STDP)
We used a biophysical model of spike timing-dependent plasticity of excitatory synapses of the network. This rule operated constantly on the weights of the excitatory synapses during simulations. Synaptic weights evolved according to a first-order dynamic (Shouval et al., 2002; Delord et al., ) under the control of intra-synaptic calcium (Graupner and Brunel, ) through
where the plastic modifications of the synapses, i.e., the phosphorylation and dephosphorylation processes of the synaptic receptor channels, depended on a kinase (e.g., PKC type) and a phosphatase (e.g., calcineurin type) whose allosteric activation was dependent on calcium. Here, Kmax represents the maximum reaction rate of the kinase, Pmax that of the phosphatase, KCa and PCa the calcium half-activation concentration, Ca the synaptic calcium concentration and nH is the Hill's coefficient. The term t-LTP, kinase-related, was independent of synaptic weight (“additive” t-LTP) while t-LTD, phosphatase-related, was weight-proportional (“multiplicative” t-LTD), consistent with the literature (Bi and Poo, ; van Rossum et al., 2000). This model of STDP is extremely simple, but a detailed implementation would be prohibitive in an RNN of the order of a thousand neurons. There was no term related to the auto-phosphorylation of CaMKII present in many models to implement a form of molecular memory, because on one hand it is not actually involved in the maintenance of memory of synaptic modifications (Chen et al., ), and on the other hand memory is ensured here by the dynamics of kinase and phosphatase at low calcium concentration (Delord et al., ).
The time dependence of the APs (Bi and Poo, ; He et al., ) came from calcium dynamics, according to the model of Graupner and Brunel (). In this model, synaptic calcium followed
where the total calcium concentration takes into account pre- and post-synaptic calcium contributions.
Pre-synaptic spiking mediated calcium dynamics followed
where the first term corresponds to calcium extrusion/buffering with time constant τCa and the second term to voltage-dependent calcium channels (VDCC)-mediated calcium entry due to pre-synaptic spiking, with Capre the amplitude of calcium entering at each AP of the presynaptic neuron, t(i) the times of the pre-synaptic APs, and D a delay modeling the time required for the activation of AMPA channels, the depolarizing rise of the associated excitatory post-synaptic potential (EPSP) and the subsequent opening of VDCC that induces this calcium entry.
Post-synaptic spiking-mediated calcium dynamics evolved according to
and modeled extrusion/buffering (first-term) as well as calcium entries due to post-synaptic, back-propagated spiking from the post-synaptic soma along the dendritic tree to the synapse, opening VDCC (central term) and NMDA channels (right term). ξPrePost is an interaction coefficient and t(j) corresponds to the AP time of the post-synaptic neuron. NMDA activation is non-linear and depends on the product of a pre- and a post-synaptic term, representing the dependence of NMDA channel openings on the associative conjunction of pre-synaptic glutamate and post-synaptic depolarization, which releases the magnesium blockade of NMDA channels.
Synaptic Scaling
Synaptic weights were subjected to a homeostatic form of synaptic normalization, present in the cortex (Turrigiano et al., 1998; Wang and Gao, 2012; Sweatt, 2016), which was modeled in a simplified, multiplicative and instantaneous form (Zenke et al., 2013), following at each time step
This procedure ensured that the sum of the incoming weights on a post-synaptic neuron remained constant despite the plastic modifications due to STDP.
Estimation of the Time Constant of STDP With Synaptic Scaling
Without synaptic scaling, ẇij = ẇSTDP = K(Ca)−P(Ca)w. However, synaptic scaling plays an important role in the slow decay of weights, so to study the time constant of this decay we needed to incorporate the effect of synaptic scaling. Considering n weights of average value μw incoming upon a post-synaptic neuron, where a proportion p of weights undergo STDP of value ẇSTDP at time step t followed by scaling, then for a given weight w within the proportion p,
so that after algebra, one obtains
Passing to the limit Δt → 0, one finds:
i.e.
To find an estimate of the time constant of plasticity, linearization aroundμw gives
so that
Theoretical Dependences Under Asynchronous Irregular Dynamics
The steady-state theoretical concentration of calcium in individual synapses was obtained from fixed-points of CaPre and CaPost, which yielded
which was used to determine STDP modification rates
and to determine the time constant for plasticity, in the case of the network asynchronous irregular regime at low frequency, where p = 1, i.e.
Weights Within and Outside the Engram
Initial excitatory weights (before the 1 h simulation) were convolved with a centered normalized Gaussian function (σ = 5 neurons). Convolved weights with values above 0.1 (times pE→E = 0.35 to take into account inexistent weights) were considered within the engram, the other weights were considered outside the engram. Both weight populations were kept constant and their evolution was studied across time (see Figures 6, 7).
Trajectory Replay Detection
In order to detect coherent propagating activity pulse packets along the synaptic pathway, we convolved spiking activity across time and neurons with centered normalized Gaussian functions (σ = 30 ms and σ ~ 10 neurons). Neurons were considered “active” when at least 40% of the convolved frequencies which include them (>5% of normalized Gaussian function maximum) are above 12.5Hz. We considered the emergence of an activity packet when it contained more than 20 neurons.
Spiking Irregularity
To capture spiking irregularity, we quantified the CV (coefficient of variation), CV2 and Lv (time-local variation) of the inter-spike interval (ISI) distribution of the spiking trains of neurons in the network (Compte, ; Shinomoto et al., 2005) according to
where CV = CV2 = Lv = 1 for a homogeneous Poisson spike train and = 0 for a perfectly regular spike train where all ISI are equal. CV stands around 1 to 2 in vivo (Compte, ; Shinomoto et al., 2005), representing the global variability of an entire ISI sequence, but is sensitive to firing rate fluctuations. CV2 and Lv stand around 0.25 to 1.25 and 0 to 2, respectively in vivo (Compte, ; Shinomoto et al., 2005), evaluating the ISI variability locally in order to be less sensitive to firing rate fluctuations. The CV was calculated on every ISI across neurons, while the CV2 and Lv were calculated for each excitatory neuron and averaged across the whole population.
Spiking Synchrony
Three measures of synchrony were adopted, a synchrony measure S (Golomb et al., ), pairwise correlation coefficient averaged over all pairs of excitatory neurons < ρ> (Tchumatchenko et al., 2010), and Fano factor F. The first two were calculated on the estimated instantaneous neural frequency f (Gaussian convolution of spikes, σ = 30ms), while the last was calculated on the population sum of spike counts s, following
These measures equal , < ρ> = 0 and F = 1 for perfectly asynchronous network activity, and S = < ρ> = 1 while F increases for perfectly synchronous network activity.
Procedures and Parameters
Models were simulated and explored using custom developed code (MATLAB) and were numerically integrated using the forward Euler method with time-step Δt = 0.5ms in network models. Unless indicated in the text, standard parameter values were as following. Concerning the network architecture, N = 605 neurons, nE = 484 neurons, nI = 121 neurons, pE→E = 0.35, pE→I = 0.2056, pI→E = 0.22, pI→I = 0.25, μw = 0.03, σw = 0.02. Concerning the Integrate-and-Fire neural properties, C = 1 μF.cm−2, θ = −52 mV, Vrest = −67 mV, ΔtAP = 3 ms. Concerning currents, , VL = −70 mV, Δtsyn = 0.5 ms, , , , , VAMPA = VNMDA = 0 mV, VGABAA = −70 mV, VGABAB = −90 mV, [Mg2+] = 1.5 mM, τAMPA = 2.5 ms, τNMDA = 62 ms, τGABAA = 10 ms, τGABAB = 25 ms, pAMPA =pNMDA = pGABAA = pGABAB = 0.1, gRec = 0.65, gE→E = gE→I = gI→E = 1, gI→I = 0.7, pAMPA.FF ~ 0.0951. Concerning synaptic properties, , KCa = 3 μM, , PCa = 2 μM, nH = 4, Ca0 = 0.1 μM, τCa = 100 ms, ΔCapre = 0.02 μM, D = 10 ms, ΔCapost = 0.02 μM, .
Results
Predicting Fundamental Plastic Properties of PFC Recurrent Networks
To evaluate neural trajectory learning, memorization and replay, we studied a local prefrontal cortex (PFC) recurrent network model, with 484 excitatory and 121 inhibitory integrate and fire (IAF) neurons with topographically tuned feed-forward inputs. Synaptic connections were constrained by cortical connectivity data, following Dale's law, sparseness and log-normal weight distributions, and α-amino-3-hydroxy-5-methyl-4-isoxazolepropionic acid (AMPA) and N-methyl-D-aspartate (NMDA) excitatory and γ-aminobutyric acid (GABA-A and GABA-B) inhibitory synaptic currents (Figure 1A; see Materials and Methods). Most synaptic and neural properties, while present in PFC, are generic across cortex, such that the following results can be generalized to non-PFC cortical areas.
Figure 1
Excitatory synapses were plastic, i.e., endowed with realistic calcium dynamics (Graupner and Brunel,
Plastic modifications operated through calcium-dependent kinase-phosphatase kinetics (Delord et al.,
Most importantly, plasticity operated online—i.e., permanently, without offline learning periods—on excitatory synaptic weights, as a function of neuronal activity in the network, whether it corresponds to the spontaneous, asynchronous and irregular (AI) activity of the network, the activity evoked by the feed-forward currents during the input presentation of an example trajectory, or the replay activity after learning (see below). Both kinase-mediated long-term spike timing-dependent potentiation (t-LTP) and phosphatase-mediated long-term spike timing-dependent depression (t-LTD) increased non-linearly with pre- and post-synaptic spiking frequency, due to the allosteric activation of enzymes by calcium (Figure 1C). However, they differed in that kinase-mediated t-LTP was independent of synaptic weight (additive or hard-bounded) while phosphatase-mediated t-LTD was weight-proportional (multiplicative or soft-bounded), consistent with the literature (Bi and Poo,
Stability of Network AI Dynamics Under Synaptic Plasticity
A potential issue of synaptic plasticity in network models remains its sensitivity to spontaneous activity. Hence, before testing the possible role of STDP in trajectory learning and replay, we first studied the effect of STDP on the spontaneous regime, with the aim of verifying that network activity remained stable over the long term and that neurons always discharged in the AI regime. Indeed, Hebbian or post-Hebbian rules of the STDP type, by modifying the matrix of synaptic weights, may lead to saturation of neuronal activity and a collapse of the complexity of the dynamics, from initially AI chaotic activity characteristic of the waking state (Destexhe et al.,
Figure 2

Stability of spontaneous irregular asynchronous (AI) network dynamics under synaptic plasticity. (A1–A3) Membrane potential of network neurons during 3 s of spontaneous AI regime in the absence of plasticity (A1), after 1 h of plasticity (A2) and after full convergence of synaptic weights due to plasticity (A3). The same initial random connectivity matrix is used for simulations in (A1–A3). Spikes indicated by black dots. Full convergence of the synaptic matrix was obtained by simulating the networks with very fast kinetic constants. (B1–B3) Synaptic weights between excitatory neurons of the network at the end of each of simulations presented in A1–A3. (C) Convergence of synaptic weights toward the mean weight of their post-synaptic neuron as a function of time, due to synaptic scaling normalization (black curves, see Results). Time evolution of the mean (red curve) and standard deviation (blue curve) of synaptic weights. For sake of clarity, only a random selection of synapses is shown. The mean is constant and the standard deviation decreases with time, due to scaling. (D). Average excitatory neural spiking frequency (D1) and irregularity (D3) and excitatory population synchrony (D2) quantifiers, as a function of time, for five different simulations of the network with different realizations of the initial random synaptic matrix. Dots on the right indicate values obtained from network simulations after full convergence of synaptic weights. Shaded areas represent 95% confidence intervals of the mean.
Simulations showed that the spontaneous activity of the network was identical without plasticity (Figure 2A1), after 1 h in the presence of plasticity (Figure 2A2) and after full convergence (Figure 2A3) of weight matrix dynamics. This observation is consistent with the absence of changes in the connectivity matrix in the presence of STDP, even after 1 h of simulation (Figures 2B2,B3), compared to the condition without STDP (Figure 2B1). Mechanistically, the low spiking frequency of neurons resulted in moderate average elevations of calcium above its basal concentration in synapses, so that kinase and phosphatase were only very weakly activated. Therefore, weights underwent extremely slow plastic modifications where additive t-LTP (which dominated the multiplicative t-LTD at weak weights) was compensated by synaptic scaling. Due to these effects, weights converged toward the mean initial weight of their post-synaptic neuron (Figure 2C) with an apparent time constant of 2 h, close to the theoretical estimation of the time constant of plasticity (see Materials and Methods and Discussion), which predicts a time constant of 1.95 h during learning at low spiking frequencies and calcium concentrations (Ca~Ca0) in the AI regime. These steady-state values were normally distributed, with a constant mean value (due to the synaptic scaling) and a decreasing standard deviation, due to the homogenization of weights within each post-synaptic neuron (Figure 2C). Even with this more homogeneous synaptic matrix (Figure 2B3), AI dynamics were preserved (Figure 2A3). Indeed, excitatory frequency was stable (Figure 2D1), as well as markers of synchrony (Figure 2D2) and irregularity (Figure 2D3). Thus, overall, the activity regime of the network was not altered by the presence of plastic processes. Note that in PFC circuits experiencing dynamically changing feed-forward inputs, convergence of the synaptic matrix may be attenuated or even non-existent.
Learning Trajectory Engrams Under AI Dynamics
Trajectory learning during network activity has already been investigated in the theoretical literature, but either without chaotic dynamics or using biologically unrealistic learning rules (see Introduction). To test for the possibility of learning trajectories within physiologically irregular activity, we presented to the network a moving stimulus (Figure 1A, feedforward connections) that successively activated all the excitatory neurons over 1,350 ms (Figure 3B). Such a stimulation corresponds to a displacement speed of ~0.3 neurons/ms, where each excitatory neuron was stimulated for ~100 ms and discharged at ~100 Hz. This single stimulus presentation triggered neural activity much stronger than the spontaneous activity, sufficient to modify the matrix of synaptic weights. Indeed, whereas the synaptic matrix was initially formed of low random weights (Figure 3A), after presentation, the weights of synapses connecting neurons activated by the stimulus at close successive times were increased (Figure 3C). This diagonal band of increased weights formed an oriented connectivity path along stimulus-activated neurons and is referred to as the trajectory engram hereafter. Weight modifications inside and outside this trajectory engram resulted from increases due to t-LTP (Figure 3D1, ΔwLTP) and decreases due to t-LTD (Figure 3D2, ΔwLTD). Moreover, the homeostatic process of synaptic scaling, which ensures the constancy of the sum of the incoming weights of the cortical neurons, decreased the total incoming synaptic weights on post-synaptic neurons, in order to compensate for weight modifications due to STDP (Figure 3D3, ΔwScaling). In fine, STDP and scaling led together to an increase in engram weights and a slight decrease in off-engram weights (Figure 3D4, ΔwTotal; also observe the darker area in Figure 3C, compared to Figure 3A).
Figure 3

Learning a trajectory stimulus into a trajectory engram. (A). Synaptic matrix between excitatory neurons prior to stimulus presentation. (B). Membrane potential of network neurons in response to the presentation of a trajectory stimulus (stimulus in red) that successively activates all excitatory neurons over a duration of 1,350 ms. Spikes indicated by black dots. (C). Synaptic matrix between excitatory neurons after stimulus presentation. (D1–D4). Weight modifications resulting, after trajectory presentation, from t-LTP (D1), t-LTD (D2), scaling (D3), and their sum (D4). (E–H) Membrane potential (E), calcium (F), plastic rates (G) and synaptic weight dynamics (H) during the passage of the trajectory stimulus in a pair of neurons with nearby topographical tuning #102 (E1) and #112 (E2) and their reciprocal connections 102 → 112 (F1-H1) and 112 → 102 (F2-H2), and in a pair of neurons with more distant topographical tuning #102 (E3) and #202 (E4) and their reciprocal connections 102 → 202 (F3-H3) and 202 → 102 (F4-H4).
The observation, on a local scale, of the details of the processes at work for the synapses linking the neurons of the engram allowed for a better understanding of these network effects. For illustration, neurons #102 and #112, with close spatial topographical tuning, discharged one following the other with partial overlap during the stimulus (Figure 3E). At the level of the synapse between neurons #102 and #112 (102 → 112), whose orientation was that of the trajectory, the arrival of pre-synaptic action potentials (APs) was followed by that of postsynaptic APs (pre #102 then post #112 neuron, Figures 3E1,E2), which triggered a massive input of calcium via the VDCC channels and the NMDA receptor channels (Figure 3F1). Conversely, in the synapse 112 → 102, for which the sequence of arrival of the APs was reversed (pre #112 then post #102 neuron), NMDA channels did not open (see above), such that the calcium input resulted only from the VDCC channels and was thus moderate (Figure 3F2). These calcium elevations activated the kinases and phosphatases, which, respectively, phosphorylated and dephosphorylated AMPA channels, increasing (t-LTP) and decreasing (t-LTD) synaptic weights (only phosphorylated AMPA channels are functional and ensure synaptic transmission). These kinase and phosphatase activations were important for synapse 102 → 112 (Figure 3G1), but less so for the synapse 112 → 102 (Figure 3G2). For both synapses (Figures 3G1,G2), the phosphatase was more strongly activated (lower half-activation; Delord et al.,
For neurons whose receptive fields were more spatially distant, activation by the stimulus occurred at more temporally distant times (for example, neurons #102 and #202, Figures 3E3,E4). In this case, regardless of the sequence of arrival of the APs in both neurons, their succession was too distant in time to open NMDA channels, so that incoming calcium came only from the VDCC channels and was therefore low (Figures 3F3,F4). Consequently, kinase and phosphatase were weakly activated, resulting in virtually null STDP velocity (Figures 3G3,G4). Synaptic scaling (Figures 3G3,G4), induced by the increase of weights in the engram (Figures 3H1,H2), ultimately decreased synaptic weights (Figures 3H3,H4). As such, there was no learning of any trajectory between distant neurons, contrary to what happened between closer neurons.
Trajectory Replays From Learned Trajectory Engrams
In behaving animals, learnt trajectories are replayed later in appropriate behavioral conditions. In the model, we assessed whether trajectories could be replayed, the dynamics of trajectory replays and the way they affect the network connectivity compared to before they occur (Figure 4A). Trajectory replay was defined as the reactivation of neurons of the entire trajectory engram, after temporarily stimulating only initial neurons at the beginning of the engram. To assess trajectory replay in the network, we applied a stimulus of 100 ms to the first 50 neurons of the engram, 500 ms after trajectory learning was completed (Figure 4B). We found that the network was able to replay the trajectory entirely after learning (Figure 4B1). Fundamentally, the replay emerged because neurons were linked by strong synapses so that preceding neurons activated subsequent neurons in the engram, forming an oriented propagating wave (Figure 4B2).
Figure 4

Replay of learned trajectories. (A). Synaptic matrix between excitatory neurons after stimulus presentation but prior to trajectory replay. (B). Membrane potential of network neurons (B1, spikes indicated by black dots) in response to the trajectory stimulus, followed by a transient trajectory replay triggered by stimulating the start of the trajectory (neurons #1–50, stimulus in red). Membrane potential of a selected subset of neurons along the trajectory (B2, arbitrary colors). (C). Synaptic matrix between excitatory neurons after stimulus and replay. (D). Weight modifications resulting, after compared to before trajectory replay, from t-LTP (D1), t-LTD (D2), scaling (D3), and their sum (D4). (E) Recapitulation of the whole trajectory after separately learning four individual trajectory fragments (ABCD) in the forward order (E1; chunking) or backward order (E2; ordinal knowledge). Each fragment corresponds to 180 neurons. Fragments overlap over 65 neurons.
Because it activated neurons at several tens of Hz, the replay could have brought into play plastic processes at the synapses forming the engram, and, in doing so, either reinforce or diminish their weights, possibly disturbing or even destroying the engram. To evaluate these possibilities, we observed the variation of synaptic weights before and after the replay. We found that after replay, the engram was still present (Figure 4C) and its structure identical to that before replay (Figure 4A). However, when dissecting the effects at work, we found that the engram had slightly thickened during the trajectory replay, due to the combined effect of t-LTP (Figure 4D1 ΔwLTP), t-LTD (Figure 4D2 ΔwLTD) and scaling (Figure 4D3 ΔwScaling). Weights above and below the engram increased, whereas weights slightly decreased within the engram (Figure 4D4, ΔwTotal, red fringes).
Up to this point, the neural trajectory was presented as a whole. However, whole trajectories are generally not accessible directly to the PFC. Rather, PFC circuits generally encounter elementary trajectory fragments at separate points in time to produce prospective planning of future behaviors (Ito et al.,
Functional Diversity of Trajectory Replays
Neural activity during the replay was less focused than the stimulus trajectory (Figure 4B), i.e., it involved more (~90 vs. 35) neurons, spiking at a lower (~65 vs. 100 Hz) discharge frequency. The replay also unfolded at a faster speed, lasting ~750 ms—for a stimulus of 1,350 ms—so that it exhibited a temporal compression factor (tCF) of ~1.8, which is situated between fast and regular timescale replays observed in animals. Regular timescale replays operate at the timescale of behaviors they were learnt from, i.e., a few seconds (in navigation or working memory tasks, e.g.), hence typically displaying tCF~1. By contrast, fast timescale replays last several hundred ms in the awake PFC (200–1,500 ms; Jadhav et al.,
Figure 5

Functional diversity of trajectory replays. (A) Trajectory replay duration (upper left white bar) and compression factor (tCF; lower right) depend on the NMDA conductance decay time constant (τNMDA, range 30–150 ms). NMDA maximal conductance was scaled (range 0.475–1.8) so as to insure similar levels of firing frequency drive during trajectory replays. Regular (A1) and fast (A2) timescale replay are due to slower and faster NMDA dynamics. (B). Single-trial (B1, B2) and inter-trial variability (B3,B4) of firing frequency of individual neurons (B1,B3) and of the population (B2,B4) for 10 different simulations similar to the replay shown Figure 4B. Lines represent mean values, shaded regions represent 95% confidence intervals of the mean.
Globally, the model thus not only indicated that it was possible to learn trajectories online by creating synaptic engrams, thanks to the STDP-type plasticity rule. It also showed that learned trajectories were functional as a memory process, in the sense that their replay was possible and globally preserved the synaptic structure of the learned engram. Finally, the model accounted for the large functional diversity of replays observed in behaving animals, both with regard to the timescale (fast vs. regular) they exhibit, as well as to the type of coding (dynamical vs. stable) they may subserve in navigational or working memory tasks.
Stability of Network AI Dynamics in the Presence of Trajectory Engrams
After evaluating the stability of the learned trajectory in the presence of AI network activity, we asked the symmetrical question, i.e., whether the engram of a previously learned trajectory could alter the irregular features of spontaneous network dynamics. Indeed, the altered synaptic structure (which implies large weights in all neurons of the recurrent network) may induce correlated activations of neurons (e.g., partial replays) resulting in runaway activity-plasticity interactions and drifts in network activity and synaptic structure. We monitored network connectivity (Figure 6A) and activity dynamics (Figures 6B1–B3) for 1 h to assess the stability of the spontaneous AI regime in the presence of the engram. We observed that following learning of the engram, synaptic weights outside the engram (i.e., responsible for the AI dynamics) increased exponentially toward their new steady-state in a very slow manner (Figure 6A) with an apparent time constant of 1.91 h, consistent with the theoretical estimation of 1.95 h (see above). This increase resulted from the decrease of within-engram large synaptic weights via synaptic scaling (Figure 6E1, see above). Despite this slow and moderate structural reorganization, AI dynamics were preserved with stable frequency (Figure 6B1), synchrony (Figure 6B2), and irregularity (Figure 6B3). Thus, overall, both the synaptic structure outside the engram as well as the spontaneous AI regime remained stable in the presence of the engram.
Figure 6

Stability of the spontaneous AI regime in the presence of the engram. (A). Average synaptic weights outside the engram after learning for 1 h. Shaded areas represent 95% confidence intervals of the mean for 5 network simulations. (B). Networks dynamics after learning for 1 h: frequency (B1), synchrony (B2), and irregularity (B3) of excitatory neurons. Shaded areas as in (A). (C). Membrane potential of neurons in the neural network for 3 s following a replay stimulation of the 50 first neurons at 1 s (C1), 1 min (C2) or 1 h (C3) after trajectory learning. (D) Synaptic matrices between excitatory neurons of the network, at the end of the simulations presented in (C). (E) Network engram synaptic weight average (E1) as well as frequency (E2) and duration (E3) of trajectory replays during 1 h after trajectory learning. Shaded areas as in (A).
Memory of Trajectory Engrams in the Presence of Network AI Dynamics
We then studied whether the spontaneous AI activity could disrupt the engram of the learned trajectory and the possibility for trajectory replay. Indeed, the trajectory engram may be gradually erased, due to AI activity at low frequency favoring t-LTD, or even amplified, due to the activity in the trajectory engram caused by plasticity (resulting in further plasticity runaway). To do so, we assessed the timescale of potential drifts in engram connectivity and activity following learning, and of the network ability to replay the engram. Intuitively, engram erasure, runaway or stability probably depended on network dynamics after learning: spontaneous AI regime, spontaneous replays, or other forms of activity.
To address these questions, we simulated the network for 1 h after trajectory learning and recorded “snapshots” of the continuous evolution of the synaptic matrix every minute. Using these successive recorded matrices as initial conditions for independent simulations of replays, we were able to quantify network ability for trajectory replay, at different times of the evolution of the network. We found that while trajectory replay occurred in full after 1 s, activating all neurons of the trajectory (Figure 6C1), it was slightly attenuated after 1 min (last neurons spiking at lower frequency; Figure 6C2) and failed after 1 h (Figure 6C3). Observing the synaptic matrix at these three moments allowed us to understand the origin of this degradation in replay ability. Indeed, whereas after 1 min (Figure 6D2), the synaptic weights of the engram changed only a little compared to 1 s (Figure 6D1), the engram was narrowed and weights attenuated after 1 h (Figure 6D3). Such degradation of the engram was probably the cause of the failure to replay the trajectory 1 h after learning.
To more precisely monitor degradation of the trajectory engram and replay, we measured averaged engram weights as well as replay frequency and duration across time. We found that the engram weights declined exponentially with a fitted time constant of 1.91 h (Figure 6E1), very close to that predicted by the theory (1.95 h). The measures of trajectory replay decreased faster than the engram weights, with time constants of ~54 min for mean frequency during the replay (Figure 6E2) and ~13 min for replay duration (Figure 6E3). Specifically, replay of the full trajectory lasted 4 min. The degradation of trajectory replay was mainly due to progressive replay failure in the neurons located later in the trajectory engram. The faster decrease in trajectory activity, compared to the average engram weights, was probably a consequence of a cooperative mechanism of propagation in the engram: the non-linearity in NMDA current activation, requiring synergistic activation of pre- and post-synaptic neurons in the engram, rendered the propagation of activity non-linearly sensitive to decreases in engram weights.
Repeated Trajectory Replays Can Destabilize Trajectory Engrams and Replays
We have observed that a single replay of the trajectory only marginally modified the engram (Figure 4C vs. Figure 4A). However, we assessed whether replay repetitions could strengthen the engram significantly further. Such strengthening through repetition could compensate for the engram erasure due to spontaneous activity after the learning (Figure 6E1) and its functional consequence, the relatively rapid loss of replay capacity (Figures 6E2,E3). Intuitively, the partial increase in weight at the border of the trajectory engram after one replay (Figure 4D4 ΔwTotal, red fringes) could, after repeated replays, be strong enough to counteract the decrease observed outside replays during memorization (Figure 6D3, light blue fringe). To test this possibility, we repeated the replay stimulus every 3 s for 30 s after the presentation of the initial trajectory stimulus (Figure 7A). We observed, from the very first seconds, and even before we could test the effect of the protocol at larger timescales, that these successive stimuli, initially triggering correct trajectory replays, rapidly led to hyperactivity involving most of the neurons in the network (Figure 7A1). Such paroxysmal activity typically appeared via avalanche dynamics activating neurons at the end of the trajectory (a fraction of the network, therefore), which propagated to the whole network at increasingly higher discharge frequencies (up to tens of Hz). Moreover, this activity had an oscillatory component, visible on the time course of the frequency of the excitatory and inhibitory neurons (Figure 7A2). This paroxysmal activity partially erased the engram of the learned trajectory via synaptic scaling (Figure 7B), making it impossible to replay the trajectory following this seizure (see last stimulus, Figure 7A1), consistent with similar effects found in empirical observation during epileptic seizures (Hu et al.,
Figure 7

Unstable engram and network dynamics after repeated trajectory replays. (A). Membrane potential of neurons in the neural network (A1) for 30 s during which a replay stimulus is performed on the first 50 neurons every 3 s. Mean activity of excitatory (red) and inhibitory (blue) neurons (A2). (B). Matrix of synaptic weights between excitatory neurons before (left) and after (right) paroxysmal network activity.
Slow Learning Stabilizes Trajectory Engram and Replays
As the repetition of replay learning led to over-activation of the trajectory with plasticity speed parameters sufficiently fast for a single stimulus presentation to be learned and replayed, we investigated how slower STDP kinetic coefficients could prevent paroxysmal activity during stimulus presentations and replays. For this, we used smaller values of Kmax and Pmax, i.e., here, divided by a factor of 6. With these values, 4 presentations of the trajectory stimulus were necessary for increasing the engram weights enough to sustain trajectory replays (Figure 8A). After such a learning protocol, the replay of the full trajectory was possible even beyond 1 h after learning (Figure 8B), whereas replay ability lasted only a few minutes with previous parameters (Figures 6E2,E3). This increase in replay memory timescale is consistent with that of the engram time constant, which was 11.5 h (Figure 8C), of the order of its theoretical estimation ~11.7 h, i.e., it was increased by a factor 6 compared to that obtained with previous parameters (1.91 and 1.95 h, respectively Figure 6E1). Remarkably, the memory of trajectory replay was increased by a factor >20 (trajectory completely replayed at >1.4 h vs. 4 min with previous parameters), so that, relatively to the timescale of the trajectory engram, the timescale for trajectory replay was further increased by a factor 3.5. Indeed, the presentation of several stimuli recruited a thicker-tailed weight distribution, with higher probability of large weights (blue curve above the red one in ~0.05–0.125; Figure 8D) but lowered probabilities of highly-weighted synapses (blue curve with negligible probabilities above 0.15; Figure 8D), because successive trajectory stimuli simultaneously evoked progressively stronger trajectory replays, recruiting more neurons at lower frequencies (Figure 8A), therefore imprinting larger engrams. Thus, slower plasticity kinetics required a larger number of successive presentations to learn the trajectory, but ensured a more robust engram involving more synapses, resulting in a better resilience to forgetting, i.e., a better quality of learning.
Figure 8

Slower learning stabilization of the engram and network dynamics. (A). Membrane potential of the neural network in response to the presentation of 4 trajectory stimuli in the presence of slower STDP learning kinetics. (B). Membrane voltage of the neural network for 3 s following a replay stimulation on the first 50 neurons at 1 s, 1 h after trajectory learning. (C). Average weight of all engram synapses after learning for 1 h. (D). Probability distribution of the synaptic weights of the excitatory synapses after 4 presentations of the trajectory stimulus during slow learning (blue), and after one presentation of the trajectory stimulus during learning with faster (standard) parameters (red). (E) Networks dynamics after learning with slow plasticity: frequency (E1), synchrony (E2), and irregularity (E3) of excitatory neurons. Shaded areas represent 95% confidence intervals of the mean for five network simulations. (F) Membrane potential of the neural network for 38 s during which a replay stimulus is performed on 50 neurons every 3 s for 10 total repetitions (as in Figure 7A) after 4 trajectory stimuli in the presence of slower STDP learning kinetics (as in Figure 8A). (G) Minimal number of stimulus presentations required to learn a replay (red) and maximal number of replays before paroxysmal activity (black), as a function of the plasticity slowdown factor expressed in units of plasticity standard time constant (i.e., by which slowdown factor plasticity rates are divided). The number of replays until explosion is evaluated with the same weight matrix (learned at standard plasticity speed or x1 slowdown) across different plasticity speeds, for better comparison of the effect of plasticity speeds on replay. Shaded areas represent 95% confidence intervals of the mean for 10 network simulations.
Finally, we assessed whether slow plasticity with multiple stimulus presentations also preserved network dynamics. AI dynamics were preserved with stable frequency (Figure 8E1), synchrony (Figure 8E2), and irregularity (Figure 8E3). We then repeated the replay stimulus every 3 s for 30 s after the presentation of the initial trajectory stimulus, a protocol which led to paroxysmal activity when considering fast plasticity. With slower kinetics, multiple replay stimuli triggered correct trajectory replays for the whole duration of the simulation (Figure 8F). We then asked whether a threshold of plasticity speed exists above which paroxysmal activity is triggered, or, conversely, the risk of paroxysmal activity linearly scales with the ability to learn fast. To do so, we parametrically explored simulations with plasticity rate divided by a slowdown factor in the range 1–10. The minimal number of stimulus presentations required to form a strong enough engram (i.e., allowing a replay) increased slowly with slower plasticity kinetics (Figure 8G, red). In parallel, the increase in the maximal number of replays before turning network dynamics into paroxysmal activity was much larger (Figure 8G, black), so that slowing plasticity kinetics increased the physiological range allowing learning while preserving network dynamics from paroxysmal activity. Hence, plasticity slow enough to preserve healthy dynamics may constitute a key constraint on the ability to learn rapidly. Furthermore, if the product of plasticity speed with the number of stimulus presentation was constant, it would indicate a linear summation of plastic effects arising from each presentation. By contrast, the number of stimulus presentations necessary for replay was lower than the factor of plasticity slowdown (5 stimuli for 10x plasticity slowdown instead of 10 stimuli, Figure 8G). This is due to successive stimulations overlapping with replays (i.e., stimulus presentations after the first one induce replays, Figure 8A), suggesting progressive facilitation of learning at slow plasticity speeds.
Discussion
Here, we show that it is possible to learn neural trajectories (dynamical representations) using a spike timing-dependent plasticity (STDP) learning rule in local PFC circuits displaying spontaneous activity in the asynchronous irregular (AI) regime. We used a physiological model of plasticity (Delord et al.,
This model was built to reproduce functional phenomenology of the PFC (learning, replays at different timescales, dynamic or persistent coding, see below), based on biophysical constraints from the experimental literature at the molecular, cellular and network levels, rather than by artificial training. If overall architectural properties of the model are observed in the PFC, such properties are also compatible with other non-prefrontal cortices with trajectory replays, lending strength to the genericity of the current study's results. For example, the excitatory/inhibitory network balance, observed in the PFC (Shu et al., 2003; Haider et al.,
In the model, external feedforward inputs are constant, as in previous models of characteristic PFC activity (for e.g., Brunel,
Molecular Plasticity and Memory in the PFC
In the PFC, e-STDP necessitates more than the pre-post synaptic pairings used in spike-timing protocols, as long-term potentiation (t-LTP) emerges in the presence of dopaminergic or cholinergic tonic neuromodulation, or when inhibitory synaptic transmission is decreased (Couey et al.,
Stable Spontaneous AI Dynamics in the PFC in the Presence of Plasticity and Learning
Hebbian forms of plasticity (Abbott and Nelson,
Many studies show that e-STDP rules are deleterious to AI dynamics such that compensating homeostatic mechanisms are required to control neuronal activity, for e.g., a metaplastic e-STDP rule with sliding-threshold (Boustani et al.,
Learning Dynamical Representations in the PFC Under AI Dynamics
Phenomenological e-STDP models fail to learn engrams in noisy AI states because of their sensitivity to spontaneous activity. The absence of STDP weight-dependence forbids learning and induces the direct loss of engrams (Boustani et al.,
Phenomenological STDP models based on neighboring spike-doublet or spike-triplet schemes often produce side effects (either sensitivity to noisy activity, or runaway plasticity) due to the temporal bounds of the pre- and post-couplings they consider (Boustani et al.,
In feed-forward networks endowed with this STDP rule, and for conditions of spiking frequency and irregularity similar to AI activity, plastic modifications essentially depend on firing frequency rather than on the precise timing of spikes, because equivalent probabilities of encountering pre-then-post and post-then-pre spike pairs in conditions of stationary spiking essentially blurs net spike-timing effects (Graupner et al.,
The previous studies that have addressed the possibility of engram learning in recurrent networks with AI dynamics focused on static stimuli (Morrison et al., 2007; Boustani et al.,
Long-Term Memory of Dynamical Representations in the PFC Under AI Dynamics
The present study underlines the importance of slow plasticity kinetics together with repeated presentations for learning dynamic representations in PFC networks. Faster kinetics allowed one-shot learning of trajectory engrams, but extensive training could then induce paroxysmal activity during the trajectory replays that partly erased the engram, which was ultimately detrimental to the learning and replay process. This synchronous increase in neuronal activity in the model is reminiscent of epileptic seizures (Truccolo et al., 2011), which have been found to cancel out the plasticity effects of synaptic weights (Hu et al.,
Fast learning together with stable memory is considered in many synaptic plasticity models to rely on auto-phosphorylation of the calmodulin-dependent protein kinase II (CaMKII). CaMKII auto-phosphorylation is appealing because it constitutes a positive-feedback loop (inducing fast plasticity) underlying bistable dynamics (providing infinite memory of a single potentiated synaptic state). However, we did not consider CaMKII in the present model, because CamKII is not necessary to the maintenance of synaptic modifications (Chen et al.,
Here, the stability of molecular memory originated from extremely slow synaptic weight dynamics, resulting in slow exponential forgetting of the engram. Slow weight dynamics arose from activity-dependent kinase and phosphatase (aKP, Delord et al.,
The timescale of trajectory replay scaled with that of the engram. This is because replay requires a sufficiently preserved engram to emerge from synaptic interactions between neurons. However, the lifetime of trajectory replay was an order of magnitude smaller than that of the trajectory engram, because replay requires neuronal interactions that are non-linear and therefore sensitive to decreases in synaptic weights. Interestingly, the long-term degradation of trajectory replay was due to incomplete replay at the end of the trajectories learned, in a manner consistent with the primacy effect of medium-term learned sequences (Greene et al.,
Humans or animals generally learn complex navigational paths such as sensory, motor or behavioral sequences in a progressive manner. Thus, PFC circuits are often challenged with the necessity to process several parts of whole neural trajectories that are discovered as sequences of elementary parts encountered at separate points in time. Moreover, prospective processes in the PFC require recombining elementary neural trajectories into new trajectory representations serving the planning of future actions, choices or navigational paths, for e.g., during rule switching and behavioral adaptation (Ito et al.,
Multiple Functional Relevance of STDP-Based Neural Trajectories in the PFC
We found in our model that the same network, taught with the same stimulus, could generate a large range of replay duration and compression factors, including those characterizing regular (Batuev,
Besides, individual neuronal activity displayed lower firing frequency during replay compared to the activity induced by the stimulus, consistent with sparse coding of representations after learning. Firing rates of individual neurons during stimuli or delays in working memory tasks, as well as in navigation tasks, vary considerably across species and behavioral contexts, spanning two orders of magnitude from ~1 to ~ 100 Hz (Fuster and Alexander,
In the PFC, representations for executive functions and cognition can present less explicit dynamic coding schemes than regular timescale neural trajectories presented here. For instance, working memory can display intricate patterns of complex (heterogeneous but non-random) dynamic activities that can hardly be disentangled into simpler well-separate transient patterns of activity (Jun et al.,
Here, trajectory replays offer a possible unified framework that can participate to reconcile opposite views regarding the nature of information persistent vs. dynamic coding in the PFC (Constantinidis et al.,
Interestingly, we found that the population activity of trajectory replays accounted for the decreasing pattern of activity that can be observed in the PFC (Cavanagh et al.,
Statements
Data availability statement
The raw data supporting the conclusions of this article will be made available by the authors, without undue reservation.
Author contributions
MS, JV, JN, and BD developed the model. MS, JV, DM, and BD contributed to numerical simulations and their analysis. MS, JV, DM, JN, and BD wrote the article. All authors contributed to the article and approved the submitted version.
Conflict of interest
The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.
References
1
AbbottL. F.NelsonS. B. (2000). Synaptic plasticity: taming the beast. Nat. Neurosci.3, 1178–1183. 10.1038/81453
2
AbelesM.BergmanH.GatI.MeilijsonI.SeidemannE.TishbyN.et al. (1995). Cortical activity flips among quasi-stationary states. Proc. Natl. Acad. Sci. U.S.A. 92, 8616–8620. 10.1073/pnas.92.19.8616
3
BaegE. H.KimY. B.HuhK.Mook-JungI.KimH. T.JungM. W. (2003). Dynamics of population code for working memory in PFC. Neuron40, 177–188. 10.1016/S0896-6273(03)00597-X
4
BakhurinK. I.GoudarV.ShobeJ. L.ClaarL. D.BuonomanoD. V.MasmanidisS. C. (2017). Differential encoding of time by prefrontal and striatal network dynamics. J. Neurosci.37, 854–870. 10.1523/JNEUROSCI.1789-16.2016
5
BarbosaJ.SteinH.MartinezR. L.Galan-GadeaA.LiS.DalmauJ.et al. (2020). Interplay between persistent activity and activity-silent dynamics in the prefrontal cortex underlies serial biases in working memory. Nat. Neurosci.23, 1016–1024. 10.1038/s41593-020-0644-4
6
BatuevA. S. (1994). Two neuronal systems involved in short-term spatial memory in monkeys. Acta Neurobiol. Exp.54, 335–344.
7
BeaulieuC.KisvardayZ.SomogyiP.CynaderM.CoweyA. (1992). Quantitative distribution of gaba-immunopositive and -immunonegative neurons and synapses in the monkey striate cortex (area 17). Cerebral Cortex2, 295–309. 10.1093/cercor/2.4.295
8
BeggsJ. M.PlenzD. (2003). Neuronal avalanches in neocortical circuits. J. Neurosci.23, 11167–11177. 10.1523/JNEUROSCI.23-35-11167.2003
9
BenchenaneK.TiesingaP. H.BattagliaF. P. (2011). Oscillations in the prefrontal cortex: a gateway to memory and attention. Curr. Opin. Neurobiol.21, 475–485. 10.1016/j.conb.2011.01.004
10
BertschingerN.NatschlägerT. (2004). Real-time computation at the edge of chaos in recurrent neural networks. Neural Comput.16, 1413–1436. 10.1162/089976604323057443
11
BiG. Q.PooM. M. (1998). Synaptic modifications in cultured hippocampal neurons: dependence on spike timing, synaptic strength, and postsynaptic cell type. J. Neurosci.18, 10464–10472. 10.1523/JNEUROSCI.18-24-10464.1998
12
BittnerK. C.MilsteinA. D.GrienbergerC.RomaniS.MageeJ. C. (2017). Behavioral time scale synaptic plasticity underlies CA1 place fields. Science357, 1033–1036. 10.1126/science.aan3846
13
BoustaniS. E.YgerP.FrégnacY.DestexheA. (2012). Stable learning in stochastic network states. J. Neurosci.32, 194–214. 10.1523/JNEUROSCI.2496-11.2012
14
BrodyC. D.HernándezA.ZainosA.RomoR. (2003). Timing and neural encoding of somatosensory parametric working memory in macaque prefrontal cortex. Cereb Cortex13, 1196–1207. 10.1093/cercor/bhg100
15
BrunelN. (2000). Dynamics of Sparsely Connected Networks of Excitatory and Inhibitory Spiking Neurons. J Comput Neurosci8, 183–208. 10.1023/A:1008925309027
16
BrunelN.WangX. J. (2001). Effects of neuromodulation in a cortical network model of object working memory dominated by recurrent inhibition. J. Comput. Neurosci.11, 63–85. 10.1023/A:1011204814320
17
BuonomanoD. V.MaassW. (2009). State-dependent computations: spatiotemporal processing in cortical networks. Nat. Rev. Neurosci.10, 113–125. 10.1038/nrn2558
18
BuschmanT. J.MillerE. K. (2014). Goal-direction and top-down control. Philos. Transac. R. Soc. B Biol. Sci.369:20130471. 10.1098/rstb.2013.0471
19
CavanaghS. E.TowersJ. P.WallisJ. D.HuntL. T.KennerleyS. W. (2018). Reconciling persistent and dynamic hypotheses of working memory coding in prefrontal cortex. Nat. Commun.9:3498. 10.1038/s41467-018-05873-3
20
ChenG.GreengardP.YanZ. (2004). Potentiation of NMDA receptor currents by dopamine D1 receptors in prefrontal cortex. PNAS101, 2596–2600. 10.1073/pnas.0308618100
21
ChenH.-X.OtmakhovN.StrackS.ColbranR. J.LismanJ. E. (2001). Is persistent activity of calcium/calmodulin-dependent kinase required for the maintenance of LTP?J. Neurophysiol.85, 1368–1376. 10.1152/jn.2001.85.4.1368
22
ChenkovN.SprekelerH.KempterR. (2017). Memory replay in balanced recurrent networks. PLoS Comput. Biol.13:e1005359. 10.1371/journal.pcbi.1005359
23
ClopathC.BüsingL.VasilakiE.GerstnerW. (2010). Connectivity reflects coding: a model of voltage-based STDP with homeostasis. Nat. Neurosci.13, 344–352. 10.1038/nn.2479
24
ClopathC.ZieglerL.VasilakiE.BüsingL.GerstnerW. (2008). Tag-trigger-consolidation: a model of early and late long-term-potentiation and depression. PLoS Comput. Biol.4:e1000248. 10.1371/journal.pcbi.1000248
25
CompteA. (2003). Temporally irregular mnemonic persistent activity in prefrontal neurons of monkeys during a delayed response task. J. Neurophysiol.90, 3441–3454. 10.1152/jn.00949.2002
26
CompteA.BrunelN.Goldman-RakicP. S.WangX.-J. (2000). Synaptic mechanisms and network dynamics underlying spatial working memory in a cortical network model. Cereb Cortex10, 910–923. 10.1093/cercor/10.9.910
27
ConstantinidisC.FunahashiS.LeeD.MurrayJ. D.QiX.-L.WangM.et al. (2018). Persistent spiking activity underlies working memory. J. Neurosci.38, 7020–7028. 10.1523/JNEUROSCI.2486-17.2018
28
CoueyJ. J.MeredithR. M.SpijkerS.PoorthuisR. B.SmitA. B.BrussaardA. B.et al. (2007). Distributed network actions by nicotine increase the threshold for spike-timing-dependent plasticity in prefrontal cortex. Neuron54, 73–87. 10.1016/j.neuron.2007.03.006
29
CromerJ. A.RoyJ. E.MillerE. K. (2010). Representation of multiple, independent categories in the primate prefrontal cortex. Neuron66, 796–807. 10.1016/j.neuron.2010.05.005
30
DaleH. (1935). Pharmacology and nerve-endings (Walter Ernest Dixon memorial lecture): (Section of Therapeutics and Pharmacology). Proc. R. Soc. Med.28, 319–332. 10.1177/003591573502800330
31
DehaeneS.MeynielF.WacongneC.WangL.PallierC. (2015). The neural representation of sequences: from transition probabilities to algebraic patterns and linguistic trees. Neuron88, 2–19. 10.1016/j.neuron.2015.09.019
32
DelordB.BerryH.GuigonE.GenetS. (2007). A new principle for information storage in an enzymatic pathway model. PLoS Comput. Biol.3:e124. 10.1371/journal.pcbi.0030124
33
DelordB.KlaassenA. J.BurnodY.CostalatR.GuigonE. (1997). Bistable behaviour in a neocortical neurone model. Neuroreport8, 1019–1023. 10.1097/00001756-199703030-00040
34
DembrowN.JohnstonD. (2014). Subcircuit-specific neuromodulation in the prefrontal cortex. Front. Neural Circuits8:54. 10.3389/fncir.2014.00054
35
DestexheA.RudolphM.ParéD. (2003). The high-conductance state of neocortical neurons in vivo. Nat. Rev. Neurosci.4, 739–751. 10.1038/nrn1198
36
DruckmannS.ChklovskiiD. B. (2012). Neuronal circuits underlying persistent representations despite time varying activity. Curr. Biol.22, 2095–2103. 10.1016/j.cub.2012.08.058
37
DudaiY. (2012). The restless engram: consolidations never end. Annu. Rev. Neurosci.35, 227–247. 10.1146/annurev-neuro-062111-150500
38
DurstewitzD.SeamansJ. K.SejnowskiT. J. (2000). Neurocomputational models of working memory. Nat. Neurosci.3, 1184–1191. 10.1038/81460
39
EckerA. S.BerensP.KelirisG. A.BethgeM.LogothetisN. K.ToliasA. S. (2010). Decorrelated neuronal firing in cortical microcircuits. Science327, 584–587. 10.1126/science.1179867
40
EllwoodI. T.PatelT.WadiaV.LeeA. T.LiptakA. T.BenderK. J.et al. (2017). Tonic or phasic stimulation of dopaminergic projections to prefrontal cortex causes mice to maintain or deviate from previously learned behavioral strategies. J. Neurosci.37, 8315–8329. 10.1523/JNEUROSCI.1221-17.2017
41
EnelP.WallisJ. D.RichE. L. (2020). Stable and dynamic representations of value in the prefrontal cortex. Elife9:e54313. 10.7554/eLife.54313.sa2
42
EnokiR.HuY.HamiltonD.FineA. (2009). Expression of long-term plasticity at individual synapses in hippocampus is graded, bidirectional, and mainly presynaptic: optical quantal analysis. Neuron62, 242–253. 10.1016/j.neuron.2009.02.026
43
EustonD. R.GruberA. J.McNaughtonB. L. (2012). The role of medial prefrontal cortex in memory and decision making. Neuron76, 1057–1070. 10.1016/j.neuron.2012.12.002
44
EustonD. R.TatsunoM.McNaughtonB. L. (2007). Fast-forward playback of recent memory sequences in prefrontal cortex during sleep. Science318, 1147–1150. 10.1126/science.1148979
45
FieteI. R.SennW.WangC. Z. H.HahnloserR. H. R. (2010). Spike-time-dependent plasticity and heterosynaptic competition organize networks to produce long scale-free sequences of neural activity. Neuron65, 563–576. 10.1016/j.neuron.2010.02.003
46
FujiiN.GraybielA. M. (2003). Representation of action sequence boundaries by macaque prefrontal cortical neurons. Science301, 1246–1249. 10.1126/science.1086872
47
FujisawaS.AmarasinghamA.HarrisonM. T.BuzsákiG. (2008). Behavior-dependent short-term assembly dynamics in the medial prefrontal cortex. Nat. Neurosci.11, 823–833. 10.1038/nn.2134
48
FunahashiS.BruceC. J.Goldman-RakicP. S. (1989). Mnemonic coding of visual space in the monkey's dorsolateral prefrontal cortex. J. Neurophysiol.61, 331–349. 10.1152/jn.1989.61.2.331
49
FusiS.DrewP. J.AbbottL. F. (2005). Cascade models of synaptically stored memories. Neuron45, 599–611. 10.1016/j.neuron.2005.02.001
50
FusterJ. M.AlexanderG. E. (1971). Neuron activity related to short-term memory. Science173, 652–654. 10.1126/science.173.3997.652
51
GallistelC. R.MatzelL. D. (2013). The neuroscience of learning: beyond the Hebbian synapse. Annu. Rev. Psychol.64, 169–200. 10.1146/annurev-psych-113011-143807
52
Goldman-RakicP. S. (1995). Cellular basis of working memory. Neuron14, 477–485. 10.1016/0896-6273(95)90304-6
53
GolombD.HanselD.MatoG. (2001). Mechanisms of synchrony of neural activity in large networks, in Handbook of Biological Physics, Volume 4: Neuro-Informatics and Neural Modelling, eds F. Moss, and S. Gielen (Amsterdam: Elsevier Science), 887–968.
54
GotoY.YangC. R.OtaniS. (2010). Functional and dysfunctional synaptic plasticity in prefrontal cortex: roles in psychiatric disorders. Biol. Psychiatry67, 199–207. 10.1016/j.biopsych.2009.08.026
55
GraupnerM.BrunelN. (2012). Calcium-based plasticity model explains sensitivity of synaptic changes to spike pattern, rate, and dendritic location. PNAS109, 3991–3996. 10.1073/pnas.1109359109
56
GraupnerM.WallischP.OstojicS. (2016). Natural firing patterns imply low sensitivity of synaptic plasticity to spike timing compared with firing rate. J. Neurosci.36, 11238–11258. 10.1523/JNEUROSCI.0104-16.2016
57
GreeneA. J.PrepsciusC.LevyW. B. (2000). Primacy versus recency in a quantitative model: activity is the critical distinction. Learn. Mem.7, 48–57. 10.1101/lm.7.1.48
58
HahnG.PetermannT.HavenithM. N.YuS.SingerW.PlenzD.et al. (2010). Neuronal avalanches in spontaneous activity in vivo. J. Neurophysiol.104, 3312–3322. 10.1152/jn.00953.2009
59
HaiderB.DuqueA.HasenstaubA. R.McCormickD. A. (2006). Neocortical network activity in vivo is generated through a dynamic balance of excitation and inhibition. J. Neurosci.26, 4535–4545. 10.1523/JNEUROSCI.5297-05.2006
60
HeK.HuertasM.HongS. Z.TieX.HellJ. W.ShouvalH.et al. (2015). Distinct eligibility traces for LTP and LTD in cortical synapses. Neuron88, 528–538. 10.1016/j.neuron.2015.09.037
61
HebbD. O. (1949). The Organization of Behavior: A Neuropsycholocigal Theory. New York, NY: John Wiley & Sons, Inc.
62
HempelC. M.HartmanK. H.WangX.-J.TurrigianoG. G.NelsonS. B. (2000). Multiple forms of short-term plasticity at excitatory synapses in rat medial prefrontal cortex. J. Neurophysiol.83, 3031–3041. 10.1152/jn.2000.83.5.3031
63
HuB.SergeiK.LeiZ.ArminS. (2005). Reversal of hippocampal LTP by spontaneous seizure-like activity: role of group i mGlur and cell depolarization. J. Neurophysiol. 93, 316–336. 10.1152/jn.00172.2004
64
IsaacsonJ. S.ScanzianiM. (2011). How inhibition shapes cortical activity. Neuron72, 231–243. 10.1016/j.neuron.2011.09.027
65
ItoH. T.ZhangS.-J.WitterM. P.MoserE. I.MoserM.-B. (2015). A prefrontal–thalamo–hippocampal circuit for goal-directed spatial navigation. Nature522, 50–55. 10.1038/nature14396
66
JadhavS. P.KemereC.GermanP. W.FrankL. M. (2012). Awake hippocampal sharp-wave ripples support spatial memory. Science336, 1454–1458. 10.1126/science.1217230
67
JadhavS. P.RothschildG.RoumisD. K.FrankL. M. (2016). Coordinated excitation and inhibition of prefrontal ensembles during awake hippocampal sharp-wave ripple events. Neuron90, 113–127. 10.1016/j.neuron.2016.02.010
68
JahrC. E.StevensC. F. (1990). Voltage dependence of NMDA-activated macroscopic conductances predicted by single-channel kinetics. J. Neurosci.10, 3178–3182. 10.1523/JNEUROSCI.10-09-03178.1990
69
JensenG.AltschulD.DanlyE.TerraceH. (2013). Transfer of a serial representation between two distinct tasks by rhesus macaques. PLoS ONE8:e70285. 10.1371/journal.pone.0070285
70
JunJ. K.MillerP.HernándezA.ZainosA.LemusL.BrodyC. D.et al. (2010). Heterogenous population coding of a short-term memory and decision task. J. Neurosci.30, 916–929. 10.1523/JNEUROSCI.2062-09.2010
71
KaeferK.NardinM.BlahnaK.CsicsvariJ. (2020). Replay of behavioral sequences in the medial prefrontal cortex during rule switching. Neuron106, 154–165.e6. 10.1016/j.neuron.2020.01.015
72
KeckC.SavinC.LückeJ. (2012). Feedforward inhibition and synaptic scaling – two sides of the same coin?PLoS Comput. Biol.8:e1002432. 10.1371/journal.pcbi.1002432
73
KeckT.ToyoizumiT.ChenL.DoironB.FeldmanD. E.FoxK.et al. (2017). Integrating Hebbian and homeostatic plasticity: the current state of the field and future research directions. Philos. Transac. R. Soc. B Biol. Sci.372:20160158. 10.1098/rstb.2016.0158
74
KlampflS.MaassW. (2013). Emergence of dynamic memory traces in cortical microcircuit models through STDP. J. Neurosci.33, 11515–11529. 10.1523/JNEUROSCI.5044-12.2013
75
La CameraG.FontaniniA.MazzucatoL. (2019). Cortical computations via metastable activity. Curr. Opin. Neurobiol.58, 37–45. 10.1016/j.conb.2019.06.007
76
LajeR.BuonomanoD. V. (2013). Robust timing and motor patterns by taming chaos in recurrent neural networks. Nat. Neurosci.16, 925–933. 10.1038/nn.3405
77
LengyelI.VossK.CammarotaM.BradshawK.BrentV.MurphyK. P. S. J.et al. (2004). Autonomous activity of CaMKII is only transiently increased following the induction of long-term potentiation in the rat hippocampus. Eur. J. Neurosci.20, 3063–3072. 10.1111/j.1460-9568.2004.03748.x
78
Litwin-KumarA.DoironB. (2014). Formation and maintenance of neuronal assemblies through synaptic plasticity. Nat. Commun.5:5319. 10.1038/ncomms6319
79
LiuJ. K.BuonomanoD. V. (2009). Embedding multiple trajectories in simulated recurrent neural networks in a self-organizing manner. J. Neurosci.29, 13172–13181. 10.1523/JNEUROSCI.2358-09.2009
80
LondonM.RothA.BeerenL.HäusserM.LathamP. E. (2010). Sensitivity to perturbations in vivo implies high noise and suggests rate coding in cortex. Nature466, 123–127. 10.1038/nature09086
81
LundqvistM.HermanP.MillerE. K. (2018). Working Memory: delay activity, yes! Persistent activity? Maybe not. J. Neurosci.38, 7013–7019. 10.1523/JNEUROSCI.2485-17.2018
82
LutzuS.CastilloP. E. (2021). Modulation of NMDA receptors by g-protein-coupled receptors: role in synaptic transmission, plasticity and beyond. Neuroscience456, 27–42. 10.1016/j.neuroscience.2020.02.019
83
MalenkaR. C.LancasterB.ZuckerR. S. (1992). Temporal limits on the rise in postsynaptic calcium required for the induction of long-term potentiation. Neuron9, 121–128. 10.1016/0896-6273(92)90227-5
84
ManninenT.HituriK.Hellgren KotaleskiJ.BlackwellK. T.LinneM.-L. (2010). Postsynaptic signal transduction models for long-term potentiation and depression. Front. Comput. Neurosci.4:152. 10.3389/fncom.2010.00152
85
ManteV.SussilloD.ShenoyK. V.NewsomeW. T. (2013). Context-dependent computation by recurrent dynamics in prefrontal cortex. Nature503, 78–84. 10.1038/nature12742
86
MarkowitzD. A.CurtisC. E.PesaranB. (2015). Multiple component networks support working memory in prefrontal cortex. PNAS112, 11084–11089. 10.1073/pnas.1504172112
87
MarkramH.GerstnerW.SjöströmP. J. (2012). Spike-timing-dependent plasticity: a comprehensive overview. Front. Synaptic Neurosci.4:2. 10.3389/fnsyn.2012.00002
88
MashhooriA.HashemniaS.McNaughtonB. L.EustonD. R.GruberA. J. (2018). Rat anterior cingulate cortex recalls features of remote reward locations after disfavoured reinforcements. Elife7:e29793. 10.7554/eLife.29793
89
MeadorK. J. (2007). The basic science of memory as it applies to epilepsy. Epilepsia48, 23–25. 10.1111/j.1528-1167.2007.01396.x
90
MeyersE. M.FreedmanD. J.KreimanG.MillerE. K.PoggioT. (2008). Dynamic population coding of category information in inferior temporal and prefrontal cortex. J. Neurophysiol.100, 1407–1419. 10.1152/jn.90248.2008
91
MoberlyA. H.SchreckM.BhattaraiJ. P.ZweifelL. S.LuoW.MaM. (2018). Olfactory inputs modulate respiration-related rhythmic activity in the prefrontal cortex and freezing behavior. Nat. Commun.9:1528. 10.1038/s41467-018-03988-1
92
MongilloG.BarakO.TsodyksM. (2008). Synaptic theory of working memory. Science319, 1543–1546. 10.1126/science.1150769
93
MontgomeryJ. M.MadisonD. V. (2002). State-dependent heterogeneity in synaptic depression between pyramidal cell pairs. Neuron33, 765–777. 10.1016/S0896-6273(02)00606-2
94
MorrisonA.AertsenA.DiesmannM. (2007). Spike-timing-dependent plasticity in balanced random networks. Neural Comput.19, 1437–1467. 10.1162/neco.2007.19.6.1437
95
MurrayJ. D.BernacchiaA.RoyN. A.ConstantinidisC.RomoR.WangX.-J. (2017). Stable population coding for working memory coexists with heterogeneous neural dynamics in prefrontal cortex. PNAS114, 394–399. 10.1073/pnas.1619449114
96
NakajimaM.SchmittL. I.HalassaM. M. (2019). Prefrontal cortex regulates sensory filtering through a basal ganglia-to-thalamus pathway. Neuron103, 445–458.e10. 10.1016/j.neuron.2019.05.026
97
NaudéJ.CessacB.BerryH.DelordB. (2013). Effects of cellular homeostatic intrinsic plasticity on dynamical and computational properties of biological recurrent neural networks. J. Neurosci.33, 15032–15043. 10.1523/JNEUROSCI.0870-13.2013
98
OnnS.-P.WangX.-B. (2005). Differential modulation of anterior cingulate cortical activity by afferents from ventral tegmental area and mediodorsal thalamus. Eur. J. Neurosci.21, 2975–2992. 10.1111/j.1460-9568.2005.04122.x
99
OnnS.-P.WangX.-B.LinM.GraceA. A. (2006). Dopamine D1 and D4 receptor subtypes differentially modulate recurrent excitatory synapses in prefrontal cortical pyramidal neurons. Neuropsychopharmacology31, 318–338. 10.1038/sj.npp.1300829
100
OstlundS. B.WinterbauerN. E.BalleineB. W. (2009). Evidence of action sequence chunking in goal-directed instrumental conditioning and its dependence on the dorsomedial prefrontal cortex. J. Neurosci.29, 8280–8287. 10.1523/JNEUROSCI.1176-09.2009
101
PapouinT.DunphyJ.TolmanM.FoleyJ. C.HaydonP. G. (2017). Astrocytic control of synaptic function. Philos. Transac. R. Soc. B Biol. Sci.372:20160154. 10.1098/rstb.2016.0154
102
PasseckerJ.MikusN.Malagon-VinaH.AnnerP.DimidschsteinJ.FishellG.et al. (2019). Activity of prefrontal neurons predict future choices during gambling. Neuron101, 152–164.e7. 10.1016/j.neuron.2018.10.050
103
PasupathyA.MillerE. K. (2005). Different time courses of learning-related activity in the prefrontal cortex and striatum. Nature433, 873–876. 10.1038/nature03287
104
PatonJ. J.BuonomanoD. V. (2018). The neural basis of timing: distributed mechanisms for diverse functions. Neuron98, 687–705. 10.1016/j.neuron.2018.03.045
105
PeyracheA.KhamassiM.BenchenaneK.WienerS. I.BattagliaF. P. (2009). Replay of rule-learning related neural patterns in the prefrontal cortex during sleep. Nat. Neurosci.12, 919–926. 10.1038/nn.2337
106
PfeifferB. E.FosterD. J. (2013). Hippocampal place-cell sequences depict future paths to remembered goals. Nature497, 74–79. 10.1038/nature12112
107
PouilleF.Marin-BurginA.AdesnikH.AtallahB. V.ScanzianiM. (2009). Input normalization by global feedforward inhibition expands cortical dynamic range. Nat. Neurosci.12, 1577–1585. 10.1038/nn.2441
108
RayeC. L.JohnsonM. K.MitchellK. J.GreeneE. J.JohnsonM. R. (2007). Refreshing: a minimal executive function. Cortex43, 135–145. 10.1016/S0010-9452(08)70451-9
109
RenartA.RochaJ.de la BarthoP.HollenderL.PargaN.ReyesA.et al. (2010). The asynchronous state in cortical circuits. Science327, 587–590. 10.1126/science.1179850
110
RodriguezG.SarazinM.ClementeA.HoldenS.PazJ. T.DelordB. (2018). Conditional bistability, a generic cellular mnemonic mechanism for robust and flexible working memory computations. J. Neurosci.38, 5209–5219. 10.1523/JNEUROSCI.1992-17.2017
111
RomoR.BrodyC. D.HernándezA.LemusL. (1999). Neuronal correlates of parametric working memory in the prefrontal cortex. Nature399, 470–473. 10.1038/20939
112
RuanH.SaurT.YaoW.-D. (2014). Dopamine-enabled anti-Hebbian timing-dependent plasticity in prefrontal circuitry. Front. Neural Circuits8:38. 10.3389/fncir.2014.00038
113
SchmittL. I.WimmerR. D.NakajimaM.HappM.MofakhamS.HalassaM. M. (2017). Thalamic amplification of cortical connectivity sustains attentional control. Nature545, 219–223. 10.1038/nature22073
114
SeidemannE.MeilijsonI.AbelesM.BergmanH.VaadiaE. (1996). Simultaneously recorded single units in the frontal cortex go through sequences of discrete and stable states in monkeys performing a delayed localization task. J. Neurosci.16, 752–768. 10.1523/JNEUROSCI.16-02-00752.1996
115
ShafiM.ZhouY.QuintanaJ.ChowC.FusterJ.BodnerM. (2007). Variability in neuronal activity in primate cortex during working memory tasks. Neuroscience146, 1082–1108. 10.1016/j.neuroscience.2006.12.072
116
ShenoyK. V.SahaniM.ChurchlandM. M. (2013). Cortical control of arm movements: a dynamical systems perspective. Annu. Rev. Neurosci.36, 337–359. 10.1146/annurev-neuro-062111-150509
117
ShimaK.IsodaM.MushiakeH.TanjiJ. (2007). Categorization of behavioural sequences in the prefrontal cortex. Nature445, 315–318. 10.1038/nature05470
118
ShinJ. D.TangW.JadhavS. P. (2019). Dynamics of awake hippocampal-prefrontal replay for spatial learning and memory-guided decision making. Neuron104, 1110–1125.e7. 10.1016/j.neuron.2019.09.012
119
ShinomotoS.MiyazakiY.TamuraH.FujitaI. (2005). Regional and laminar differences in in vivo firing patterns of primate cortical neurons. J. Neurophysiol.94, 567–575. 10.1152/jn.00896.2004
120
ShinomotoS.ShimaK.TanjiJ. (2003). Differences in spiking patterns among cortical neurons. Neural Comput.15, 2823–2842. 10.1162/089976603322518759
121
ShouvalH. Z.BearM. F.CooperL. N. (2002). A unified model of NMDA receptor-dependent bidirectional synaptic plasticity. PNAS99, 10831–10836. 10.1073/pnas.152343099
122
ShuY.HasenstaubA.McCormickD. A. (2003). Turning on and off recurrent balanced cortical activity. Nature423, 288–293. 10.1038/nature01616
123
SiapasA. G.LubenovE. V.WilsonM. A. (2005). Prefrontal phase locking to hippocampal theta oscillations. Neuron46, 141–151. 10.1016/j.neuron.2005.02.028
124
SiriB. B.QuoyM.DelordB.CessacB.BerryH. (2007). Effects of Hebbian learning on the dynamics and structure of random networks with inhibitory and excitatory neurons. J. Physiol. Paris101, 136–148. 10.1016/j.jphysparis.2007.10.003
125
SongS.SjöströmP. J.ReiglM.NelsonS.ChklovskiiD. B. (2005). Highly nonrandom features of synaptic connectivity in local cortical circuits. PLoS Biol.3:e68. 10.1371/journal.pbio.0030068
126
SpaakE.WatanabeK.FunahashiS.StokesM. G. (2017). Stable and dynamic coding for working memory in primate prefrontal cortex. J. Neurosci.37, 6503–6516. 10.1523/JNEUROSCI.3364-16.2017
127
SreenivasanK. K.CurtisC. E.D'EspositoM. (2014). Revisiting the role of persistent neural activity during working memory. Trends Cogn. Sci.18, 82–89. 10.1016/j.tics.2013.12.001
128
StokesM. G. (2015). ‘Activity-silent' working memory in prefrontal cortex: a dynamic coding framework. Trends Cogn. Sci.19, 394–405. 10.1016/j.tics.2015.05.004
129
SweattJ. D. (2016). Dynamic DNA methylation controls glutamate receptor trafficking and synaptic scaling. J. Neurochem.137, 312–330. 10.1111/jnc.13564
130
TanakaJ.HoriikeY.MatsuzakiM.MiyazakiT.Ellis-DaviesG. C. R.KasaiH. (2008). Protein synthesis and neurotrophin-dependent structural plasticity of single dendritic spines. Science319, 1683–1687. 10.1126/science.1152864
131
TchumatchenkoT.GeiselT.VolgushevM.WolfF. (2010). Signatures of synchrony in pairwise count correlations. Front. Comput. Neurosci.4:1. 10.3389/neuro.10.001.2010
132
ThomsonA. M. (2002). Synaptic connections and small circuits involving excitatory and inhibitory neurons in layers 2-5 of adult rat and cat neocortex: triple intracellular recordings and biocytin labelling in vitro. Cerebral Cortex12, 936–953. 10.1093/cercor/12.9.936
133
TiganjZ.JungM. W.KimJ.HowardM. W. (2017). Sequential firing codes for time in rodent medial prefrontal cortex. Cerebral Cortex27, 5663–5671. 10.1093/cercor/bhw336
134
TononiG.CirelliC. (2003). Sleep and synaptic homeostasis: a hypothesis. Brain Res. Bull.62, 143–150. 10.1016/j.brainresbull.2003.09.004
135
TruccoloW.DonoghueJ. A.HochbergL. R.EskandarE. N.MadsenJ. R.AndersonW. S.et al. (2011). Single-neuron dynamics in human focal epilepsy. Nat. Neurosci.14, 635–641. 10.1038/nn.2782
136
TurrigianoG. G.LeslieK. R.DesaiN. S.RutherfordL. C.NelsonS. B. (1998). Activity-dependent scaling of quantal amplitude in neocortical neurons. Nature391, 892–896. 10.1038/36103
137
van RossumM. C.BiG. Q.TurrigianoG. G. (2000). Stable Hebbian learning from spike timing-dependent plasticity. J. Neurosci.20, 8812–8821. 10.1523/JNEUROSCI.20-23-08812.2000
138
VogelsT. P.SprekelerH.ZenkeF.ClopathC.GerstnerW. (2011). Inhibitory plasticity balances excitation and inhibition in sensory pathways and memory networks. Science334, 1569–1573. 10.1126/science.1211095
139
WangH.-X.GaoW.-J. (2012). Prolonged exposure to NMDAR antagonist induces cell-type specific changes of glutamatergic receptors in rat prefrontal cortex. Neuropharmacology62, 1808–1822. 10.1016/j.neuropharm.2011.11.024
140
WangJ.NarainD.HosseiniE. A.JazayeriM. (2018). Flexible timing by temporal scaling of cortical responses. Nat. Neurosci.21, 102–110. 10.1038/s41593-017-0028-6
141
WangX.-J. (2001). Synaptic reverberation underlying mnemonic persistent activity. Trends Neurosci.24, 455–463. 10.1016/S0166-2236(00)01868-3
142
WangY.MarkramH.GoodmanP. H.BergerT. K.MaJ.Goldman-RakicP. S. (2006). Heterogeneity in the pyramidal network of the medial prefrontal cortex. Nat. Neurosci.9, 534–542. 10.1038/nn1670
143
WitterM. P.AmaralD. G. (2004). CHAPTER 21 - Hippocampal Formation, in The Rat Nervous System (Third Edition), ed G. Paxinos (Burlington, NJ: Academic Press), 635–704. 10.1016/B978-012547638-6/50022-5
144
WutzA.LoonisR.RoyJ. E.DonoghueJ. A.MillerE. K. (2018). Different levels of category abstraction by different dynamics in different prefrontal areas. Neuron97, 716–726.e8. 10.1016/j.neuron.2018.01.009
145
XuT.-X.YaoW.-D. (2010). D1 and D2 dopamine receptors in separate circuits cooperate to drive associative long-term potentiation in the prefrontal cortex. PNAS107, 16366–16371. 10.1073/pnas.1004108107
146
XueM.AtallahB. V.ScanzianiM. (2014). Equalizing excitation–inhibition ratios across visual cortical neurons. Nature511, 596–600. 10.1038/nature13321
147
YangS.-T.ShiY.WangQ.PengJ.-Y.LiB.-M. (2014). Neuronal representation of working memory in the medial prefrontal cortex of rats. Mol. Brain7:61. 10.1186/s13041-014-0061-2
148
YuJ. Y.LiuD. F.LobackA.GrossrubatscherI.FrankL. M. (2018). Specific hippocampal representations are linked to generalized cortical representations in memory. Nat. Commun.9:2209. 10.1038/s41467-018-04498-w
149
ZenkeF.GerstnerW.GanguliS. (2017). The temporal paradox of Hebbian learning and homeostatic plasticity. Curr. Opin. Neurobiol.43, 166–176. 10.1016/j.conb.2017.03.015
150
ZenkeF.HennequinG.GerstnerW. (2013). Synaptic plasticity in neural networks needs homeostasis with a fast rate detector. PLoS Comput. Biol.9:e1003330. 10.1371/journal.pcbi.1003330
151
ZhangW.LindenD. J. (2003). The other side of the engram: experience-driven changes in neuronal intrinsic excitability. Nat. Rev. Neurosci.4:885. 10.1038/nrn1248
152
ZielinskiM. C.ShinJ. D.JadhavS. P. (2019). Coherent coding of spatial position mediated by theta oscillations in the hippocampus and prefrontal cortex. J. Neurosci.39, 4550–4565. 10.1523/JNEUROSCI.0106-19.2019
Summary
Keywords
prefrontal cortex, neural trajectory, attractor, persistent and dynamical coding, working memory, learning, replay, asynchronous irregular state
Citation
Sarazin MXB, Victor J, Medernach D, Naudé J and Delord B (2021) Online Learning and Memory of Neural Trajectory Replays for Prefrontal Persistent and Dynamic Representations in the Irregular Asynchronous State. Front. Neural Circuits 15:648538. doi: 10.3389/fncir.2021.648538
Received
31 December 2020
Accepted
31 May 2021
Published
08 July 2021
Volume
15 - 2021
Edited by
Shintaro Funahashi, Kyoto University, Japan
Reviewed by
Shantanu P. Jadhav, Brandeis University, United States; Lukas Ian Schmitt, RIKEN Center for Brain Science (CBS), Japan
Updates

Check for updates
Copyright
© 2021 Sarazin, Victor, Medernach, Naudé and Delord.
This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.
*Correspondence: Matthieu X. B. Sarazin matthieu.sarazin@live.frBruno Delord bruno.delord@sorbonne-universite.fr
†These authors have contributed equally to this work
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.