Abstract
Interactive activation models (IAMs) simulate orthographic and phonological processes in implicit memory tasks, but they neither account for associative relations between words nor explicit memory performance. To overcome both limitations, we introduce the associative read-out model (AROM), an IAM extended by an associative layer implementing long-term associations between words. According to Hebbian learning, two words were defined as “associated” if they co-occurred significantly often in the sentences of a large corpus. In a study-test task, a greater amount of associated items in the stimulus set increased the “yes” response rates of non-learned and learned words. To model test-phase performance, the associative layer is initialized with greater activation for learned than for non-learned items. Because IAMs scale inhibitory activation changes by the initial activation, learned items gain a greater signal variability than non-learned items, irrespective of the choice of the free parameters. This explains why the slope of the z-transformed receiver-operating characteristics (z-ROCs) is lower one during recognition memory. When fitting the model to the empirical z-ROCs, it likewise predicted which word is recognized with which probability at the item-level. Since many of the strongest associates reflect semantic relations to the presented word (e.g., synonymy), the AROM merges form-based aspects of meaning representation with meaning relations between words.
Introduction
Interactive activation models (IAMs) have been used successfully to predict human word recognition performance, when the task implicitly requires retrieval of orthographic or phonological word forms from memory, such as perceptual identification, naming, lexical decision, or word stem completion (e.g., McClelland and Rumelhart, 1981; Grainger and Jacobs, ; Perry et al., 2007; Klonek et al., 2009). However, IAMs have not yet been applied to model performance in explicit memory tasks, such as the recognition of a set of studied words. Since Berry et al. () propose that the same signals are detected in implicit and explicit memory, in this paper we explored the versatility of IAMs to predict explicit memory performance. This seemed like a natural extension, given that in an implicit memory task the multiple read-out model (MROM; Grainger and Jacobs, ) already successfully predicted receiver operation characteristics (ROCs; Jacobs et al., 2003), which are crucial for the development of formal memory theories (cf. Malmberg, 2008, for a recent overview).
A distinctive strength of IAMs is that they allow item-level predictions for various dependent variables, such as “yes” response rates, mean response times, or mean amplitudes in electrophysiological responses (e.g., Spieler and Balota, 1997; Perry et al., 2007; Hofmann et al., 2008; Rey et al., 2009). IAMs are currently able to simulate effects resulting from lexical whole word representations or from smaller, sub-lexical representations during word recognition (e.g., Perry et al., 2007; cf. Ziegler and Goswami, 2005). So far, however, they neglect the fact that words are embedded into an experimental context of other meaningful words that potentially share a common learning history with the target word. Contextual between-word associations – as for instance the semantic relation of “lung” to its hypernym “organ” – were discussed as extension possibilities for connectionist models (e.g., Rumelhart and McClelland, 1982; Seidenberg and McClelland, 1989; Coltheart et al., ). However, they were never used for quantitative performance predictions. Such inter-item associations are better understood in the explicit memory literature (e.g., Roediger and McDermott, 1995; Nelson et al., 1998; Kimball et al., 2007), while item-level predictions of recognition memory performance are still lacking. Therefore, the present study aimed to keep the IAMs’ quantitative strengths of z-ROC and item-level predictions, while seeking to overcome an important weakness: predicting the impact of associative relations between words in an explicit recognition memory task.
Does associative-spreading activate “false memories”?
The probably best-known associative memory phenomenon is the so-called “false memory effect” (Deese, ; Roediger and McDermott, 1995): learning associated items (e.g., “table,” “sit,” “legs”) to a non-learned target item (e.g., “chair”) favors its erroneous recall or recognition. Moreover, when learning “chair” in the company of such associates, its “veridical recall” is more likely (Kimball et al., 2007). These experiments rely on tediously collected free association performance to define associations in subjective terms: a target is presented and participants name the first associates coming to their minds. Learning all of the most strongly associated items increases the target’s retrieval probability in a later memory experiment. However, such an experimental design takes only a small subset of the possible associations between the items of an experiment into account (Ratcliff and McKoon, 1994). The present study tested a simple co-occurrence approach allowing to consider all associations between all items (cf. Landauer and Dumais, 1997; Bullinaria and Levy, ; Griffiths et al., ; Jones and Mewhort, 2007; Andrews et al., ): two words were defined as being “associated” when they occurred significantly more often together than alone in a sample of 43 million sentences (Quasthoff et al., 2006) 1. Hebbian learning is the only assumption required for this definition: stimuli being repeatedly presented together are likely to be associated (Hebb, 1949; Rapp and Wettler, 1991).
Roediger and McDermott (1995) compared targets of which all of the most strongly associated items were learned, to targets of which no (freely) associated item occurred in the experimental context. Here, we challenged this rationale in a more fine-grained, parametric fashion. We hypothesized that the more associates occurred to a non-learned (new) target in the stimulus set, the greater is the amount of erroneous “yes” responses. Similarly, learning the most strongly associated items of an old target word should increase the tendency to freely recall it (Kimball et al., 2007). This led to the hypothesis that learned targets with more associates in the stimulus set should produce greater recognition rates. We tested these hypotheses in a study-test paradigm with the experimental factors old/new and co-occurrence level (low/high): low co-occurrence target items had less than eight significantly co-occurring items in the stimulus set, and high co-occurrence words had at least eight.
To theoretically frame these hypotheses, we extended the MROM by an associative layer (Grainger and Jacobs, ). The MROM consists of three layers of interacting processing units (Figure 1): the visual features of the target stimuli serve as input variables for the model’s feature layer. Feature units activate letter units, which in turn excite units of the orthographic word layer (McClelland and Rumelhart, 1981). In the associative read-out model (AROM), an associative unit for each item presented in the experiment was added. Since the process of word identification is necessary for recognizing it as learned, each association unit obtained an excitatory word identification signal from its corresponding orthographic word unit. The co-occurrence statistics implemented excitatory associative connections between the units in the associative layer. These linkages are assumed to reflect the pre-wired long-term structure of the human associative memory system that matured by experience with words (Hebb, 1949). When the target item is presented to the model, its association unit transiently activates all associated item units, which in turn activate the target unit. Thus, the greater the associative-spreading along the associative connections is (Collins and Loftus, ; Anderson, ), the greater is the activation “echo” from associated units back to the target item’s unit (Nelson et al., 1998). Since greater activation signals of IAMs typically predict a greater amount of “yes” responses (e.g., Grainger and Jacobs, ; Hofmann et al., 2008), the AROM allows the following hypothesis: the more associated items a target has, the larger is its associative activation. This should result in a greater amount of “yes” responses for both, new and old high co-occurrence target items. Apart from this qualitative, condition-wise prediction, we fitted the AROM to (cross-condition) ROCs, and tested whether the obtained signal strengths accounted for item-level variances.
Figure 1
Can each item’s signal be detected in an explicit memory task?
To allow for signal-detection analyses, participants in the experimental study were instructed to rate their recognition confidence on a six-point scale ranging from “sure no” (“1”) to “sure yes” (“6”). For all but the ROC analyses, “4” (“unsure yes”) to “6” counted as “yes” response. Based on these confidence ratings, the signal-detection approach (Green and Swets,
Figure 2

Distributions of the AMSSs and the resulting z-ROCs for the low (left panels) and high co-occurrence conditions (right panels). The first row displays the associative memory signal strength (AMSS) distributions transformed to smoothed probability density functions for the four experimental conditions. The second row depicts these functions transformed into cumulative “yes”-response probabilities and the five response criteria C(i) for i = 1–5. The third row shows the empirical and modeled z-ROCs.
Jacobs et al. (2003) equated model activations with signal strength to predict z-ROCs from the MROM’s activations. To adopt signal-detection theory’s assumption of greater signal strengths in old items (Green and Swets,
To predict human performance, Jacobs et al. (2003) defined signal strength as the mean activation across the first seven cycles (see also Grainger and Jacobs,
For obtaining signal strength distributions, the resulting AMSS values were transformed to functional forms for all four experimental conditions, i.e., the new and old low and high co-occurrence conditions. Since AMSS is conceptualized as the signal strength of the items, an additional source of variability of the items’ fixed signal strengths was required. Therefore, smoothed kernel density functions were applied for the transformation to functional forms, and the smoothing kernel factor κ was the only free parameter required for z-ROC generation. κ reflects the variability of the deterministic AMSS values of the items. The empirically obtained “yes” response probabilities were used as C(i; cf. Figure 2, second row). The parameters were optimized by fitting the model-generated z-ROCs to the empirical z-ROCs by minimizing the sum of the least squared errors of the slopes and intercepts of the low and high co-occurrence conditions (cf. Figure 2, third row). We then tested whether the z-ROC slopes of the participants deviated from those predicted by the model. These model tests were run for low and high co-occurrence conditions, separately.
Once the parameters were fixed, the AROM was challenged to account for item-level variance. Previous signal-detection-capable models of recognition memory (Malmberg, 2008) targeted the signal strengths of the items, but did not specify which particular word stimulus provides which signal strength (e.g., Glanzer et al.,
Simulation Methods: The AROM and Its Predictions
The feature, letter, and word layers, as well as their connections remained largely2 unaltered compared to the AROM’s predecessors (McClelland and Rumelhart, 1981; Grainger and Jacobs,
Word identification signals from the orthographic word layer activated the associative units (cf. Figure 1). Each Associative word unit x in cycle c obtained input activation Ax(c) by excitatory connections from the corresponding Orthographic word unit activation of the last cycle [Ox(c−1)]. Please note that only if the activation Ox(c−1) crosses the activation threshold (of zero), excitation or inhibition take place. To indicate that a variable must fulfill the logical condition of being positive, we use the subscript ± . Thus, for instance Ox+ (c−1) = Ox(c−1) if it is positive [Ox(c−1) > 0], otherwise Ox+ (c−1) = 0. The excitation from the orthographic to the associative layer was scaled by the free excitatory parameter αoa:
If a word y was significantly co-occurring to the word x (i.e., x ∧ y), an excitatory associative connection was added. It was quantified by log10-transformed χ2 values of within-sentence co-occurrence statistics crossing the significance threshold of χ2 = 6.63 (P < 0.01; Dunning,
This function implements the active spreading of associative activation. Moreover, all activated words [e.g., Ay(c−1)] inhibited each other [e.g., Ax(c−1)] by an amount scaled by a free parameter γaa:
According to this architecture, the AROM predicts greater activations, and thus a greater amount of “yes” responses for target items with a greater amount of associated items in the stimulus set. Thus, the summed net change nx(c) of each association unit is a function of the amount of e excitatory units (i.e., the number of significantly co-occurring items), and a function of all N neighbor units that cross the activation threshold. These are inhibiting the respective unit. Thus, the summed net change can be written as:
For simulating episodic memory traces, the resting levels ρ were constrained to be larger for old [ρold] than for new items [ρnew, i.e., ρold > ρnew]. Resting levels are referred to by cycle c = 0, i.e., Ax(0) = ρold for all old items, and Ax(0) = ρnew for all new units. All units cross the activation threshold at resting level, and thus inhibit and excite other units. As each unit can be connected to each other unit, but the association of a unit to itself is set to zero, 25,440 associations between the units are possible (1602–160 items). 1,402 of these associations (i.e., significant co-occurrences) were apparent. Most of the connections are inhibitory and thus a negative net inhibition nx(0) follows from cycle 0, i.e., n(x) is negative. To obtain non-linear dynamics, with a minimum activation m = −1, IAMs weight net inhibitory changes by the associative activation of the unit Ax(c−1) itself. The applied associative activation formula is thus finally (cf. McClelland and Rumelhart, 1981, p. 381, formula 3 and 4 while decay is zero in this case):
Even when the net changes would be of equal variance across nold(0) and nnew(0), Aold(1) = ρold − nold(0)*(ρold − m) will produce a greater variance across all old target item units than Anew(1) = ρnew − nnew(0)*(ρnew − m) for all new items, because the resting level scales the activation change by multiplication. As a consequence, activation variability must be greater when greater resting levels are assumed in learned items.
Formally, the memory signal strength of the target item’s unit is defined as AMSS (cf. Introduction):
We tested whether the old item units reveal a greater variance across these first seven cycles (AMSS) than new item units in the following parameter space, using step-sizes of 0.01: αoa from 0.04 to 0.09, αaa from 0.03 to 0.08, γaa from 0.03 to 0.08, ρnew from 0.01 to 0.05, and ρold from 0.06 to 0.1. This resulted in 5400 parameter sets.
To fit the simulated to the empirical z-ROCs, we transformed the AMSSs of the four experimental conditions into smoothed density functions, using the smoothing kernel factor κ as free parameter (cf. Bowman and Azzalini,
Experimental Methods: Testing the AROM’s Predictions
Participants
The participants were 30 native German speakers (17 female, mean age: 29.5, SE: 2.39, range: 16–60) without known reading disorders. They had normal or corrected-to-normal sight, and were paid for participation or received course credits.
Corpus
Word frequency and co-occurrence measures were taken from the German corpus of the “Wortschatz” project (status: December 20063; Quasthoff et al., 2006). They are based on 800 million tokens and 43 million sentences. The corpus is largely composed of online newspapers (1992–2006). To allow the AROM’s testability in 69 languages, corpus-size independent word frequency class measures of this cross-linguistic project were used. Therefore, a power function relates the frequency of each word to the most frequent word, i.e., “der” [the] is 2class more frequent than the given word (cf. Adelman and Brown,
Stimuli
Each cell in the 2 × 2 design (factors: old/new and co-occurrence) contained 40 nouns. Stimuli of the high co-occurrence conditions had at least eight significantly co-occurring neighbors in the stimulus set, and low co-occurrence stimuli less than eight. To rule out biased effects due to confounding variables (all Fs < 0.5, cf. Table 1), we controlled for emotional valence, arousal, imageability number of orthographic neighbors (Coltheart et al.,
Table 1
| Factors: old/new Co-occurrence | New | Old | ||
|---|---|---|---|---|
| Low | High | Low | High | |
| Number of stimuli | 40 | 40 | 40 | 40 |
| Number of significantly co-occurring items in the stimulus set | 3.85 (1.70) | 13.90 (4.73) | 3.80 (1.70) | 13.85 (4.04) |
| Emotional valence | 0.11 (1.31) | 0.03 (1.23) | 0.03 (1.09) | 0.00 (1.34) |
| Imageability | 3.89 (1.22) | 3.96 (1.29) | 3.94 (1.20) | 4.01 (1.43) |
| Arousal | 2.95 (0.53) | 2.90 (0.58) | 2.87 (0.54) | 2.98 (0.64) |
| Word frequency class | 11.65 (0.70) | 11.65 (0.70) | 11.68 (0.47) | 11.57 (0.71) |
| Number of orthographic neighbors | 1.57 (2.26) | 1.32 (1.81) | 1.57 (2.14) | 1.70 (2.29) |
| Bigram frequency | 17520 (106001) | 17556 (9263) | 16489 (9149) | 16648 (8546) |
| Number of letters | 6.08 (1.07) | 6.12 (1.14) | 6.22 (1.10) | 5.97 (1.35) |
Displays the means (SD) of the manipulated and controlled variables of the target stimuli in the four experimental conditions. Emotional valence ranges from −3 to +3. Imageability and arousal range from 0 to 5.
Procedure
Eighty old words were presented in the study phase and all 160 words in the test phase. Participants were instructed to judge how confident they are that a target stimulus was presented in the previous study phase (“yes”), or not (“no”). Participants were informed that they receive feedback about their error scores after the test phase. Performance data were acquired using a computer mouse. Stimuli were presented by Presentation 9.9 software (Neurobehavioral Systems Inc., Canada). To familiarize the participants with the task, five practice items were presented before the study and the test phase, respectively.
Study phase
Each trial began with a fixation cross remaining on the screen for 500 ms followed by a stimulus presented for 1500 ms. Five hash marks (“#####”) appeared until a mouse button was pressed. To avoid primacy and recency effects, three filler items were presented before and after the critical stimuli.
Test phase
A fixation cross was presented for 500 ms. Target stimuli were presented for 1500 ms, followed by a blank screen of 1500 ms. A rating scale appeared on the screen, and the participants judged their recognition–confidence via mouse clicks on a six-point scale ranging from “1” (“sure no”) to “6” (“sure yes”). For a random number of participants, this assignment was reversed during the experiment, but not for the analyses. Participants were instructed to use all confidence judgments approximately equally often. A blank screen of 500 ms was presented before the next trial started with a new fixation cross. None of the filler and practice items had any significantly co-occurring item in the critical stimulus set.
Experimental and Modeling Results
A 2 × 2 repeated measures ANOVA on the percentage of “yes” responses revealed a significant old/new effect [F(1,29) = 167.77, P < 0.001, ηp2 = 0.85]. Old items produced more “yes” responses. Moreover, a significant effect of co-occurrence was obtained [F(1,29) = 21.91, P < 0.001, ηp2 = 0.43], but no significant interaction (F < 1). The planned comparison revealed that high co-occurrence new stimuli produced a greater “yes” response rate (M = 0.2; SE = 0.02) than low co-occurrence new stimuli [M = 0.13; SE = 0.02; t(29) = 3.80, P < 0.001]. High co-occurrence old stimuli (M = 0.76; SE = 0.04) produced more “yes” responses than low co-occurrence old stimuli (M = 0.69; SE = 0.03; t(29) = 3.02, P < 0.005]. Averaged across participants, the z-ROC slopes were 0.66 in the low co-occurrence condition and 0.70 in the high co-occurrence condition. Figure 2 displays the empirical z-ROCs with slopes smaller than 1.
All 5400 parameter sets used for optimal parameter estimation revealed a greater AMSS variance for old than for new target items (Figure 3). The least squared differences between the modeled and the empirically obtained z-ROCs for low and high co-occurrence items were obtained for the parameters of αoa = 0.09, αaa = 0.03, γaa = 0.04, ρnew = 0.05, ρold = 0.07, and κ = 10.09. The parameters were fixed at these values. Simulated z-ROC slopes were 0.75 and 0.77 for the low and high co-occurrence conditions, respectively. For both co-occurrence conditions, the modeled z-scores for the five criteria of the new and old items were tested for their capability to predict the 10 z-scores empirically obtained (cf. Jacobs et al., 2003): the model’s z-scores accounted for 99.61% of the variance of the low co-occurrence data [F(1,9) = 2044.85; P < 0.001; RMSD = 0.07], and 99.05% of the high co-occurrence z-scores [F(1,9) = 834.45, P < 0.001, RMSD = 0.10]. The behaviorally obtained z-ROC slopes of the low and high co-occurrence conditions for the individual participants did not differ significantly from the z-ROC slopes predicted by the model [t(29) = 0.17; t(29) = 1.55; Ps > 0.1]4. The new target items AMSS scores accounted for 14.32% of the variance of the “yes” response probabilities [F(1,79) = 13.04], and the old targets for 10.45% [F(1,79) = 9.10; Ps < 0.001; RMSDs = 0.08].
Figure 3

AMSSs of the word units predicting the empirical “yes” response probabilities of each of the new and old items.
General Discussion
The present study provides two novel IAM features: relying on Hebb’s proposal that the repeated co-exposure of stimuli leads to their “association” (Hebb, 1949), we correctly predicted that a higher amount of associations in the stimulus set lead to higher proportions of “yes” responses to non-learned and learned items in recognition memory for words. Second, we extended a localist connectionist word recognition model by an associative layer, and showed that this AROM predicts recognition memory performance from the core cross-condition level of ROCs down to the fine-grained item-level.
The effect in non-learned items is related to “false memories” but goes beyond Roediger and McDermott (1995) seminal work. The false memory effect consisted of the comparison of target items from which either all of the most strongly associated, or no (freely) associated items were learned. The present study revealed similar effects in a recognition memory task. However, defining associations by co-occurrence statistics allowed for taking all associations between all items of the stimulus set into account. Still, when a target item contained more associations in the stimulus set, a significant effect of co-occurrence indicated more “yes” responses for new words.
For learned target items, we discovered that many associations boost recognition memory performance, as indicated by a co-occurrence effect for old words. Since the present stimulus set was carefully controlled for all kinds of psycholinguistic single-word features, we suggest that both, the co-occurrence effects to new and old items, can be attributed to the manipulation of the amount of associations of a target item.
To account for both of these findings, the co-occurrence statistics were embedded into an associative activation-spreading network (Collins and Loftus,
For extending the scope of IAMs to explicit memory processing, we implemented memory traces from the study-phase presentation according to signal-detection theory (Berry et al.,
Apart from these proof-of-concept explanations, the present study aimed at fitting the actual z-ROCs to low and high co-occurrence words by the AROM. Predicting z-ROCs from the AMSS of the items involves a modeling challenge well-known in recognition memory research (Gillund and Shiffrin,
In addition to the z-ROC parameter, two free parameters were necessary for the (old and new item units’) resting levels, and three scaled the excitation from the orthographic word to the associative layer, as well as excitation and inhibition within the associative layer. After fitting these free parameters, the empirically obtained z-ROC slopes and z-scores did not deviate from those predicted by the model (Figure 2).
The unequal variance signal-detection model just begged the answer to the question of why the z-ROC slope is smaller than one, by assuming greater signal strength variances of old items (Green and Swets,
Although the earliest associative activation-spreading models did not discuss false and veridical recognition, they can predict these effects (Roediger et al., 2001; cf. Anderson,
The AROM is novel in that it provides quantitative associative-spreading predictions for recognition memory performance. Neither any other spreading-activation model, nor any recognition memory model simulates word recognition with the same depth as the AROM: it predicts which word is recognized with which probability depending on the amount of its associates. The more associated items are in the stimulus set for a non-learned or learned target item, the larger is the probability to classify it as old. Thereby, the false memory logic is elevated to a level capable of making item-level predictions. For veridical memory of old items, the AROM’s item-level performance is somewhat lower than for the false memories in new items (see also Figure 2, lower right panel). This potentially results from the need to consider a second source of information for the prediction of old items (e.g., Yonelinas, 1994; DeCarlo,
The item-level variances were predicted by associative cross-trial excitation from the associative context of the experiment to the target items. The more associated items are presented before the target, the larger is its unit’s activation in cycle 1. Moreover, learned items still have a larger activation than non-learned items at this cycle. Starting at cycle 4 the visual input of the feature layer reaches the associative layer, and the identification of the stimulus begins to cue the associative memory layer (Gillund and Shiffrin,
As is evident from Figure 4, the associates to a target item reflect intuitively valid associations. Moreover, the AROM can simulate semantic relations in the narrowest sense of the term, as e.g., the associate [lung] is a hyponym of the target [organ]; [vice] can be considered as a hypernym of [egoism]; [virtue] can be conceived as the antonym of [vice]; and [wedding] and [marriage] are (partial) synonyms. Why do the most strongly co-activated associates often reveal a semantic relation to the target? We think that this is because semantically related items are very likely to share many associates with the target. Consider each common associate as a common associative feature of two words (cf. e.g., Shiffrin and Steyvers, 1997). When two words share many of these contextual properties, the probability increases that they are not only related in an associative sense, because they co-occur together, but also that they will be likely to reveal a semantic relation. Moreover, as the associative layer receives input from the orthographic layer, the unique identity of a word is not only defined by its associations, but also by its orthographic form-properties.
Figure 4

Exemplary association functions for respectively one target of the four stimulus conditions. Upper panels represent old target items and the lower depict new ones. Left and right panels display low and high co-occurrence stimuli, respectively. The y-axes indicate the associative layer activations. The x-axes indicate simulation cycles. Cycle 0 activations depict resting levels for learned (ρold = 0.07) and non-learned stimuli (ρnew = 0.05), implementing all events before the present trial. When associative excitation and inhibition generated the cycle 1 activations, these define the state of the cognitive system when a test–trial starts. The target items (and their association functions) are shown in (boxed) red (lines), old associates in green (solid lines), and the new associates in blue (dashed lines). Though the AMSSs as predictor variable in Figures 2 and 3 reflect mean activations across cycles 1–7, the strongest associates at cycle 50 are shown for face validity purposes (activations > ρnew).
Does the orthographic layer of the AROM account for additional item-level variance? The simplifying assumption of no top-down feedback from the associative layer allowed us to test whether not only associative, but also mere orthographic similarities between items can be accounted for by the MROM within the AROM (cf. e.g., Grainger and Jacobs,
Finally, form-properties and semantic-associative properties of words were both proposed to be crucial for morpheme representations (e.g., Devlin et al.,
Form-constituents of meaning have been extensively modeled using distributed representations (e.g., Plaut et al., 1996; Harm and Seidenberg, 2004; see also Grainger and Ziegler,
Conclusion
This study introduces the AROM as a model capturing explicit memory performance for IAMs. Associative-spreading-activation inserted into the MROM can account for cross-condition z-ROCs, condition-wise effects of associations in new and old items, and item-level performance. Given that many words most strongly associated by the model reflect semantic relations (e.g., hyponomy), the AROM should be a convenient tool for future investigations of semantic effects in word recognition, particularly also for the tasks IAMs were originally designed for: implicit memory tasks.
Statements
Acknowledgments
We like to thank Steffen Fritzemeier for stimulus selection, as well as Jens Eisermann, Annette Kinder, Rich Shiffrin, Andy Yonelinas, Joe Ziegler, and the reviewers for stimulating discussions. This work was supported by the Deutsche Forschungsgemeinschaft (research unit “Conflicts as signals in cognitive systems,” Jacobs, JA 823/4-2) and the cluster of excellence “Languages of Emotion” to the Freie Universität Berlin.
Conflict of interest
The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.
Footnotes
1.^http://corpora.uni-leipzig.de/
2.^The original IAM was designed for four-letter stimuli (McClelland and Rumelhart, 1981), and the present stimulus set contained three- to eight-letter stimuli. Therefore, blank letters were used (cf. Coltheart et al.,
3.^http://corpora.informatik.uni-leipzig.de/
4.^Zero “yes” responses were treated as one “yes” response, and only “yes” responses were treated as all but one “yes” responses, to allow for z-transformation.
References
1
AdelmanJ. S.BrownG. D. A. (2008). Modelling lexical decision: the form of frequency and diversity effects. Psychol. Rev.115, 214–229.10.1037/0033-295X.115.1.214
2
AndersonJ. R. (1983). A spreading activation theory of semantic processing. J. Verb. Learn. Verb. Behav.22, 261–295.10.1016/S0022-5371(83)90201-3
3
AndrewsM.ViglioccoG.VinsonD. (2009). Integrating experiential and distributional data to learn semantic representations. Psychol. Rev.116, 463–498.10.1037/a0016242
4
BaayenR. H.PiepenbrockR.GulikersL. (1995). The CELEX Lexical Database (CD-ROM). Linguistic Data Consortium, University of Pennsylvania, Philadelphia, PA.
5
BerryD.ShanksD. R.HensonR. N. A. (2008). A unitary signal-detection model of implicit and explicit memory. Trends Cogn. Sci.12, 367–373.10.1016/j.tics.2008.06.005
6
BogaczR.UsherM.ZhangJ.McClellandJ. L. (2007). Extending a biologically inspired model of choice: multi-alternatives, nonlinearity, and value-based multidimensional choice. Philos. Trans. R. Soc. B362, 1655–1670.10.1098/rstb.2007.2059
7
BowmanA. W.AzzaliniA. (1997). Applied Smoothing Techniques for Data Analysis. New York: Oxford University Press.
8
BraunM.JacobsA. M.HahneA.RickerB.HofmannM.HutzlerF. (2006). Model-generated lexical activity predicts graded ERP amplitudes in lexical decision. Brain Res.1073–1074, 431–439.
9
BullinariaJ. A.LevyJ. P. (2007). Extracting semantic representations from word co-occurrence statistics: a computational study. Behav. Res. Methods39, 510–526.10.3758/BF03193020
10
CollinsA.QuillianM. (1969). Retrieval time from semantic memory. J. Verb. Learn. Verb. Behav.8, 240–247.10.1016/S0022-5371(69)80069-1
11
CollinsA. M.LoftusE. F. (1975). A spreading-activation theory of semantic processing. Psychol. Rev.82, 407–428.10.1037/0033-295X.82.6.407
12
ColtheartM.DavelaarE.JonassonJ. T.BesnerD. (1977). “Access to the internal lexicon,” in Attention and Performance, ed. DornicS. (Hillsdale, NJ: Erlbaum), 535–555.
13
ColtheartM.RastleK.PerryC.LangdonR.ZieglerJ. (2001). DRC: a dual route cascaded model of visual word recognition and reading aloud. Psychol. Rev.108, 204–256.10.1037/0033-295X.108.1.204
14
ConradM.TammS.CarreirasM.JacobsA. M. (2010). Simulating syllable frequency effects within an interactive activation framework. J. Cogn. Psychol.22, 861–893.10.1080/09541440903356777
15
CorteseM. J.KhannaM. M.HackerS. (2010). Recognition memory for 2,578 monosyllabic words. Memory18, 595–609.10.1080/09658211.2010.493892
16
DankerJ. F.GunnP.AndersonJ. R. (2008). A rational account of memory predicts left prefrontal activation during controlled retrieval. Cereb. Cortex18, 2674–2684.10.1093/cercor/bhn027
17
DeCarloL. T. (2002). Signal detection theory with finite mixture distributions: theoretical developments with applications to recognition memory. Psychol. Rev.109, 710–721.10.1037/0033-295X.109.4.710
18
DeeseJ. (1959). On the prediction of occurrence of particular verbal intrusions in immediate recall. J. Exp. Psychol.58, 17–22.10.1037/h0046671
19
DevlinJ.JamisonH.MatthewsP.GonnermanL. (2004). Morphology and the internal structure of words. Proc. Natl. Acad. Sci. U.S.A.101, 14984.10.1073/pnas.0400023101
20
DunningT. (1993). Accurate methods for the statistics of surprise and coincidence. Comput. Linguist.19, 61–74
21
FirthJ. R. (1957). “A synopsis of linguistic theory 1930–1955,” in Studies in Linguistic Analysis, edited by Philological Society (Oxford: Blackwell), 1–32.
22
GillundG.ShiffrinR. (1984). A retrieval model for both recognition and recall. Psychol. Rev.91, 1–67.10.1037/0033-295X.91.1.1
23
GlanzerM.AdamsJ.IversonG.KimK. (1993). The regularities of recognition memory. Psychol. Rev.100, 546–567.10.1037/0033-295X.100.3.546
24
GlanzerM.KimK.HilfordA.AdamsJ. K. (1999). Slope of the receiver-operating characteristic in recognition memory. J. Exp. Psychol. Learn. Mem. Cogn.25, 500–513.10.1037/0278-7393.25.2.500
25
GraingerJ.JacobsA. M. (1996). Orthographic processing in visual word recognition: a multiple read-out model. Psychol. Rev.103, 518–565.10.1037/0033-295X.103.3.518
26
GraingerJ.JacobsA. M. (1998). “On localist connectionism and psychological science,” in Localist Connectionist Approaches to Human Cognition, eds GraingerJ.JacobsA. M. (Mahwah, NJ: Lawrence Erlbaum Associates Inc.), 1–38.
27
GraingerJ.ZieglerJ. (2011). A dual-route approach to orthographic processing. Front. Psychol.2:54.10.3389/fpsyg.2011.00054
28
GreenD. M.SwetsJ. A. (1966). Signal Detection Theory and Psychophysics. New York: Wiley.
29
GriffithsT.SteyversM.TenenbaumJ. (2007). Topics in semantic representation. Psychol. Rev.114, 211–244.10.1037/0033-295X.114.2.211
30
GrossbergS. (1978). “A theory of visual coding, memory, and development,” in Formal Theories of Visual Perception, eds LaurensE.LeeuwenbergH. F.BuffartJ. M. (New York: Wiley), 7–26.
31
GrossbergS. (1987). Competitive learning: from interactive activation to adaptive resonance. Cogn. Sci.11, 23–63.10.1111/j.1551-6708.1987.tb00862.x
32
HarmM.SeidenbergM. (2004). Computing the meanings of words in reading: cooperative division of labor between visual and phonological processes. Psychol. Rev.111, 662–720.10.1037/0033-295X.111.3.662
33
HebbD. (1949). The Organization of Behavior. New York: Wiley.
34
HofmannM. J.KuchinkeL.TammS.VõM. L. H.JacobsA. M. (2009). Affective processing within of a second: high arousal is necessary for early facilitative processing of negative but not positive words. Cogn. Affect. Behav. Neurosci.9, 389–397.10.3758/9.4.389
35
HofmannM. J.StennekenP.ConradM.JacobsA. M. (2007). Sublexical frequency measures for orthographic and phonological units in German. Behav. Res. Methods39, 620–629.10.3758/BF03193034
36
HofmannM. J.TammS.BraunM. M.DambacherM.HahneA.JacobsA. M. (2008). Conflict monitoring engages the mediofrontal cortex during nonword processing. Neuroreport19, 25–29.10.1097/WNR.0b013e3282f3b134
37
JacobsA. M.GrafR.KinderA. (2003). Receiver operating characteristics in the lexical decision task: evidence for a simple signal-detection process simulated by the multiple read-out model. J. Exp. Psychol. Learn. Mem. Cogn.29, 481–488.10.1037/0278-7393.29.3.481
38
JacobsA. M.GraingerJ. (1994). Models of visual word recognition – sampling the state of the art. J. Exp. Psychol. Hum. Percept. Perform.20, 1311–1334.10.1037/0096-1523.20.6.1311
39
JonesM. N.MewhortD. J. K. (2007). Representing word meaning and order information in a composite holographic lexicon. Psychol. Rev.114, 1–37.10.1037/0033-295X.114.1.1
40
KimballD. R.SmithT. A.KahanaM. J. (2007). The fSAM Model of False Recall. Psychol. Rev.114, 954–993.10.1037/0033-295X.114.4.954
41
KlonekF.TammS.HofmannM.JacobsA. M. (2009). Does familiarity or conflict account for performance in the word-stem completion task? Evidence from behavioural and event-related-potential data. Psychol. Res.73, 871–882.10.1007/s00426-008-0189-8
42
KuchinkeL.JacobsA. M.VõM. L.-H.ConradM.GrubichC.HerrmannM. (2006). Modulation of prefrontal activation by emotional words in recognition memory. Neuroreport17, 1037–1041.10.1097/01.wnr.0000221838.27879.fe
43
LandauerT. K.DumaisS. T. (1997). A solution to plato’s problem: the latent semantic analysis theory of acquisition, induction, and representation of knowledge. Psychol. Rev.104, 211–240.10.1037/0033-295X.104.2.211
44
MalmbergK. (2008). Recognition memory: a review of the critical findings and an integrated theory for relating them. Cogn. Psychol.57, 335–384.10.1016/j.cogpsych.2008.02.004
45
MassonM. E. J. (1991). “A distributed memory model of context effects in word identfication,” in Basic Processes in Reading, eds BesnerD.HumphreysG. W. (Hillsdale, NJ: Lawrence Erlbaum Associates), 233–263.
46
MassonM. E. J. (1995). A distributed memory model of semantic priming. J. Exp. Psychol. Learn. Mem. Cogn.21, 3–23.10.1037/0278-7393.21.1.3
47
McClellandJ. L. (1993). “Toward a theory of information processing in graded, random, and interactive networks,” in Attention and Performance XIV: Synergies in Experimental Psychology, Artificial Intelligence, and Cognitive Neuroscience, eds MeyerD. E.KornblumS. (Cambridge: MIT Press), 655–688.
48
McClellandJ. L.ChappellM. (1998). Familarity breeds differentiation: a subjective-likelihood approach to the effects of experience in recognition memory. Psychol. Rev.105, 724–760.10.1037/0033-295X.105.4.734-760
49
McClellandJ. L.RumelhartD. E. (1981). An interactive activation model of context effects in letter perception: I. An account of basic findings. Psychol. Rev.88, 375–407.10.1037/0033-295X.88.5.375
50
MortonJ. (1964). A preliminary functional model for language behavior. Int. Audiol.3, 216–225.10.3109/05384916409074089
51
MortonJ. (1969). Interaction of information in word recognition. Psychol. Rev.76, 165–178.10.1037/h0027366
52
MurdockB. (1997). Context and mediators in a theory of distributed associative memory (TODAM2). Psychol. Rev.104, 839–862.10.1037/0033-295X.104.4.839
53
NelsonD.McKinneyV.GeeN.JanczuraG. (1998). Interpreting the influence of implicitly activated memories on recall and recognition. Psychol. Rev.105, 299–324.10.1037/0033-295X.105.2.299
54
NormanD. A. (1968). Toward a theory of memory and attention. Psychol. Rev.75, 522–536.10.1037/h0026699
55
O’ReillyR. (1998). Six principles for biologically based computational models of cortical cognition. Trends Cogn. Sci.2, 455–462.10.1016/S1364-6613(98)01241-8
56
PageM. (2000). Connectionist modelling in psychology: a localist manifesto. Behav. Brain Sci.23, 443–512.10.1017/S0140525X00533354
57
PerryC.ZieglerJ. C.ZorziM. (2007). Nested incremental modeling in the development of computational theories: The CDP+ model of reading aloud. Psychol. Rev.114, 273–315.10.1037/0033-295X.114.2.273
58
PlautD. C.McClellandJ. L.SeidenbergM. S.PattersonK. (1996). Understanding normal and impaired word reading: computational principles in quasi-regular domains. Psychol. Rev.103, 56–115.10.1037/0033-295X.103.1.56
59
QuasthoffU.RichterM.BiemannC. (2006). “Corpus portal for search in monolingual corpora,”Proceedings of LREC-06, Genoa.
60
QuillianR. (1967). Word concepts: a theory and simualation of some basic semantic capabilities. Behav. Sci.12, 410–430.10.1002/bs.3830120511
61
RappR.WettlerM. (1991). Prediction of Free Word Associations Based on Hebbian Learning. Paper presented at the Proceedings of the International Joint Conference on Neural Networks, Singapore.
62
RatcliffR.McKoonG. (1994). Retrieving information from memory: spreading-activation theories versus compound-cue theories. Psychol. Rev.99, 518–535.10.1037/0033-295X.99.3.518
63
RatcliffR.SheuC.-F.GronlundS.-D. (1992). Testing global memory models using ROC curves. Psychol. Rev.99, 518–535.10.1037/0033-295X.99.3.518
64
ReyA.DufauS.MassolS.GraingerJ. (2009). Testing computational models of letter perception with item-level ERPs. Cogn. Neuropsychol.26, 7–22.10.1080/09541440802176300
65
RoedigerH.McDermottK. (1995). Creating false memories: Remembering words not presented in lists. J. Exp. Psychol. Learn. Mem. Cogn.21, 803–814.10.1037/0278-7393.21.4.803
66
RoedigerH. L.BalotaD. A.WatsonJ. M. (2001). “Spreading activaiton and arousal of false memory,” in The Nature of Remembering: Essays in honor of Robert G. Crowder, Science Conference Series, eds RoedigerH. L.NairneJ. S.NeathI.SurprenantA. M. (Washington, DC: American Psychological Association), 95–115.
67
RogersT.McClellandJ. (2008). Precise of semantic cognition: a parallel distributed processing approach. Behav. Brain Sci.31, 689–714.10.1017/S0140525X0800589X
68
RumelhartD.ToddP. (1993). “Learning and connectionist representations,” in Attention and Performance XIV: Synergies in Experimental Psychology, Artificial Intelligence, and Cognitive Neuroscience, eds MeyerD. E.KornblumS. (Cambridge, MA: MIT Press), 3–30.
69
RumelhartD. E.McClellandJ. L. (1982). An interactive activation model of context effects in letter perception: II. The contextual enhancement effect and some tests and extensions of the model. Psychol. Rev.89, 60–94.10.1037/0033-295X.89.1.60
70
SeidenbergM.McClellandJ. (1989). A distributed, developmental model of word recognition and naming. Psychol. Rev.96, 523–568.10.1037/0033-295X.96.4.523
71
ShiffrinR.SteyversM. (1997). A model for recognition memory: REM-retrieving effectively from memory. Psychol. Bull. Rev.4, 145–166.10.3758/BF03209391
72
SpielerD. H.BalotaD. A. (1997). Bringing computational models of word naming down to the item level. Psychol. Sci.8, 411–416.10.1111/j.1467-9280.1997.tb00453.x
73
SquireL.WixtedJ.ClarkR. (2007). Recognition memory and the medial temporal lobe: a new perspective. Nat. Rev. Neurosci.8, 872–883.10.1038/nrn2154
74
SteyversM.GriffithsT.DennisS. (2006). Probabilistic inference in human semantic memory. Trends Cogn. Sci.10, 327–334.10.1016/j.tics.2006.05.005
75
Thompson-SchillS. L.BotvinickM. M. (2006). Resolving conflict: a response to Martin and Cheng (2006). Psychol. Bull. Rev.13, 402–408.10.3758/BF03193860
76
VõM. L.-H.ConradM.KuchinkeL.UrtonK.HofmannM. J.JacobsA. M. (2009). The Berlin affective word list reloaded (BAWLR). Behav. Res. Methods41, 534–538.10.3758/BRM.41.2.534
77
WixtedJ. (2007). Dual-process theory and signal-detection theory of recognition memory. Psychol. Rev.114, 152–176.10.1037/0033-295X.114.1.203
78
YonelinasA. P. (1994). Receiver-operating characteristics in recognition memory: evidence for a dual-process model. J. Exp. Psychol. Learn. Mem. Cogn.20, 1341–1354.10.1037/0278-7393.20.6.1341
79
YonelinasA. P.OttenL. J.ShawK. N.RuggM. D. (2005). Separating the brain regions involved in recollection and familiarity in recognition memory. J. Neurosci.25, 3002–3008.10.1523/JNEUROSCI.5295-04.2005
80
ZieglerJ.GoswamiU. (2005). Reading acquisition, developmental dyslexia, and skilled reading across languages: a psycholinguistic grain size theory. Psychol. Bull. Rev.131, 3–29.10.1037/0033-2909.131.1.3
Summary
Keywords
unequal variance signal-detection model, associative-spreading-activation, veridical and false memory, contextual, semantic, interactive activation model, multiple read-out model, co-occurrence statistics
Citation
Hofmann MJ, Kuchinke L, Biemann C, Tamm S and Jacobs AM (2011) Remembering Words in Context as Predicted by an Associative Read-Out Model. Front. Psychology 2:252. doi: 10.3389/fpsyg.2011.00252
Received
18 May 2011
Accepted
12 September 2011
Published
04 October 2011
Volume
2 - 2011
Edited by
Jonathan Grainger, CNRS, France
Reviewed by
Thomas Hannagan, Aix-Marseille University/CNRS, France; Jeff Bowers, University of Bristol, UK; Conrad Perry, Swinburne University of Technology, Australia
Copyright
© 2011 Hofmann, Kuchinke, Biemann, Tamm and Jacobs.
This is an open-access article subject to a non-exclusive license between the authors and Frontiers Media SA, which permits use, distribution and reproduction in other forums, provided the original authors and source are credited and other Frontiers conditions are complied with.
*Correspondence: Markus J. Hofmann, Neurocognitive Psychology, Department of Psychology, Room JK 27/239, Habelschwerdter Allee 45, 14195 Berlin, Germany. e-mail: mhof@zedat.fu-berlin.de
This article was submitted to Frontiers in Language Sciences, a specialty of Frontiers in Psychology.
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.