<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.2 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.2/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.2" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">2767-0279</journal-id>
<journal-title-group>
<journal-title>Glossa Psycholinguistics</journal-title>
</journal-title-group>
<issn pub-type="epub">2767-0279</issn>
<publisher>
<publisher-name>eScholarship Publishing</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5070/G6011.50683</article-id>
<article-categories>
<subj-group>
<subject>Brief article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Morphological variation and priming in Hungarian spontaneous dialogue</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<contrib-id contrib-id-type="orcid">https://orcid.org/0000-0001-7896-4801</contrib-id>
<name>
<surname>R&#225;cz</surname>
<given-names>P&#233;ter</given-names>
</name>
<email>racz.peter.marton@ttk.bme.hu</email>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>Budapest University of Technology and Economics</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2026-06-05">
<day>05</day>
<month>06</month>
<year>2026</year>
</pub-date>
<pub-date pub-type="collection">
<year>2026</year>
</pub-date>
<volume>5</volume>
<issue>1</issue>
<elocation-id>11</elocation-id>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2026 The Author(s)</copyright-statement>
<copyright-year>2026</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="https://glossapsycholinguistics.journalpub.escholarship.org/articles/10.5070/G6011.50683/"/>
<abstract>
<p>I examine morphological priming in spontaneous dialogue by tracking the alternation between the long (-j&#225;l/-j&#233;l) and the short (-j) variant of the 2<sc>sg</sc> indefinite subjunctive in Hungarian task-oriented conversations, to see whether members of dyads persist and converge in their choice of suffixed variants. I combine conversational data from the Akaka Maptask and Budapest Games corpora with distributional estimates from the Hungarian Webcorpus, resulting in a dataset of 334 subjunctive tokens. Hierarchical Bayesian generalised linear models show robust within-verb priming: following a long prime, a long target is about 3x more likely, with no comparable effect across different verbs. Speaker identity and prime&#8211;target distance did not improve model fit. Independent of priming, conversational choices track corpus baselines, indicating an additive contribution of lexical distributions. The results provide naturalistic evidence for the priming of inflected morphological variants in Hungarian, and clarify how lexical distributions scaffold variant choice in conversation.</p>
</abstract>
</article-meta>
</front>
<body>
<sec>
<title>1. Introduction</title>
<sec>
<title>1.1 Morphological priming</title>
<p>When language users encounter linguistic material, this facilitates the subsequent production of identical or related material. In experimental psycholinguistics, this facilitation effect is called <italic>priming</italic>: exposure to a prime stimulus reduces processing costs for a related target. In linguistic interaction, the same process maps onto how Speaker 1 and Speaker 2 use a given construction.</p>
<p>In conversation, multiple mechanisms affect the facilitation effect. <italic>Self-priming</italic> is the facilitation of a speaker&#8217;s own subsequent production by their own prior output. <italic>Reciprocal priming</italic> occurs when one speaker&#8217;s output facilitates the other speaker&#8217;s production of the same or a related form. <italic>Alignment</italic> and <italic>convergence</italic> refer to the broader process by which interlocutors come to use increasingly similar forms over the course of an interaction, which may arise from reciprocal priming, shared situational pressures, or both (<xref ref-type="bibr" rid="B9">Garrod &#38; Pickering, 2009</xref>; <xref ref-type="bibr" rid="B28">Pickering &#38; Garrod, 2004</xref>). The absence of a speaker-identity effect in a conversational dataset is consistent with reciprocal priming, but does not rule out parallel self-priming under shared lexical biases.</p>
<p>A similar set of distinctions applies within morphology. The <italic>lexical boost</italic> is the facilitation of production by recent activation of the same lemma or word form (<xref ref-type="bibr" rid="B27">Pickering &#38; Branigan, 1998</xref>). This is distinct from <italic>structural priming</italic>, which operates on abstract syntactic configurations independently of lexical content (<xref ref-type="bibr" rid="B5">Bock, 1986</xref>; <xref ref-type="bibr" rid="B13">Gries, 2005</xref>). Within morphological priming specifically, one can distinguish <italic>stem-level</italic> priming (facilitation of a lemma), <italic>affix-level</italic> priming (facilitation of a suffix or inflectional pattern across different stems), and <italic>whole-inflected-form</italic> priming (facilitation of a specific inflected token). Evidence for affix-level priming comes primarily from experimental work on derivational and inflectional decomposition (<xref ref-type="bibr" rid="B2">Amenta &#38; Crepaldi, 2012</xref>; <xref ref-type="bibr" rid="B22">Marslen-Wilson, 2007</xref>). In naturally occurring speech, it is difficult to distinguish affix-level from whole-form priming, because the same speaker typically re-uses the same verb in the same inflected form. Whole-form priming is often referred to with a theoretically more neutral label, <italic>persistence</italic>, in conversational data &#8211; see Szmrecsanyi (<xref ref-type="bibr" rid="B38">2006</xref>) below.</p>
<p>In psycholinguistics, lexical boost is well established, and we have clear evidence for the priming effect of specific suffixes as well (see e.g. <xref ref-type="bibr" rid="B2">Amenta &#38; Crepaldi, 2012</xref>). People access and produce stems and affixes more readily after recent exposure.</p>
<p>R&#225;cz and Luk&#225;cs (<xref ref-type="bibr" rid="B33">2024</xref>) and R&#225;cz et al. (<xref ref-type="bibr" rid="B32">2020</xref>) demonstrated alignment between language users not only in the use of stems or affixes, but in entire distributions of variation in morphological inflection. Through a series of online word-picking games, they showed that players could converge to the choices of their co-player and that the effects of this interaction persisted after the game in post-testing. The effects themselves were not limited to across-word or across-suffix priming, but rather entailed a shift in lexical distributions at the subword level: the co-player&#8217;s word choice influenced the behaviour of similar words and the strength of this effect was a function of their similarity to it.</p>
<p>In models of dialogue as joint action, convergence provides a core coordination mechanism and is a lynchpin of language variation and change (<xref ref-type="bibr" rid="B4">Beckner et al., 2009</xref>; <xref ref-type="bibr" rid="B8">Garrod &#38; Pickering, 2004</xref>, <xref ref-type="bibr" rid="B9">2009</xref>; <xref ref-type="bibr" rid="B11">Giles &#38; Coupland, 1991</xref>; <xref ref-type="bibr" rid="B28">Pickering &#38; Garrod, 2004</xref>, <xref ref-type="bibr" rid="B29">2006</xref>). Convergence has been attested in naming preferences (<xref ref-type="bibr" rid="B34">Roberts, 2010</xref>), syntax (<xref ref-type="bibr" rid="B13">Gries, 2005</xref>), and phonetics (<xref ref-type="bibr" rid="B3">Babel, 2012</xref>). In comparison, morphological convergence in natural language has received limited attention. Weiner and Labov (<xref ref-type="bibr" rid="B41">1983</xref>) found that a strong predictor of passive use in English-language interviews was the presence of another passive in the previous five utterances. Szmrecsanyi (<xref ref-type="bibr" rid="B38">2006</xref>) used a corpus-based approach to look at the English comparative, where a word-formation pattern (<italic>friendlier</italic>) alternates with an analytical construction (<italic>more friendly</italic>). He found that the choice of either variant persisted in discourse, meaning that if speakers started using one, they would be more likely to keep using it.</p>
<p>The dearth of similar research work in natural language reflects the challenges of detecting morphological convergence in interactional data. Many analytic languages, like English, provide a shallow pool of morphological variation, which is, in turn, very hard to dredge in corpora. Lab studies, like R&#225;cz et al. (<xref ref-type="bibr" rid="B32">2020</xref>), used an experimental battery that exposed participants to a rapid fire of nonwords and a framework that actively encouraged predicting the co-player&#8217;s behaviour. This allowed these studies to explore more nuanced convergence effects across lexical distributions. On the flip side, the design was divorced from natural language use, in which convergence usually operates on familiar words and its incentives and mechanisms are more implicit. The natural-language studies, like Weiner and Labov (<xref ref-type="bibr" rid="B41">1983</xref>), relied on natural language and identified patterns that were frequent enough in English to constitute a robust, compact sample.</p>
<p>This study uses a set of dyadic interactions recorded in Hungarian. Interactive tasks entail the frequent use of instructions, requests, and commands. Hungarian uses the subjunctive to express these imperative functions (<xref ref-type="bibr" rid="B39">T&#243;th, 2007</xref>). The Hungarian 2<sc>sg</sc> subjunctive suffix varies between a long and a short form. This structured variation can be tracked across interactions and offers a unique window into how lexical distributions and convergence effects shape word use together in a language with rich inflection and complex word-formation processes.</p>
</sec>
<sec>
<title>1.2 The Hungarian subjunctive</title>
<p>The Hungarian verb is marked for person, number, tense, mood, and definiteness. The subjunctive, used both in the imperative and in certain subordinate constructions (<xref ref-type="bibr" rid="B39">T&#243;th, 2007</xref>), is marked in the present tense and has a different paradigm for definite and indefinite verbs. The 2<sc>sg</sc> form varies in both paradigms, with a long and a short form. In this article, I focus on the more frequent indefinite. Example (1) shows the long and the short form.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(1)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>kelt-s&#233;l/kelt-s</p></list-item>
<list-item><p>wake-<sc>sbjv</sc>-2<sc>sg-indef</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>fel,</p></list-item>
<list-item><p><sc>part</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>ha</p></list-item>
<list-item><p>if</p></list-item>
</list>
<list list-type="word">
<list-item><p>szeptember</p></list-item>
<list-item><p>September</p></list-item>
</list>
<list list-type="word">
<list-item><p>v&#233;get</p></list-item>
<list-item><p>end.<sc>acc</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>&#233;r</p></list-item>
<list-item><p>reach-3<sc>sg-indef</sc></p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#160;&#160;&#8216;Wake me up when September ends.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The short form is an underlying -<italic>j</italic> which can assimilate to the preceding stem, while the long form adds -<italic>j&#225;l/-j&#233;l</italic>, depending on the backness of the final stem vowel, following the general patterns of Hungarian vowel harmony (<xref ref-type="bibr" rid="B36">Sipt&#225;r &#38; T&#246;rkenczy, 2000</xref>) (hence, <italic>kelts&#233;l/*keltj&#233;l</italic> wake-<sc>sbjv</sc>-2<sc>sg-indef</sc>). The long/short variation is attested, in some form, with all types of verbs, including irregular forms (see e.g. <italic>l&#233;gy/legy&#233;l</italic> be-<sc>sbjv</sc>-2<sc>sg-indef</sc>, <italic>gyere/j&#246;jj&#233;l</italic> come-<sc>sbjv</sc>-2<sc>sg-indef</sc>). This variation has been present in Hungarian since at least the 15th century (<xref ref-type="bibr" rid="B18">L&#225;szl&#243;, 2001</xref>), but no systematic, large-scale, quantitative study exists of the stylistic and social factors that condition it.</p>
</sec>
</sec>
<sec>
<title>2. Current study</title>
<p>I used a Hungarian Webcorpus and a dataset of dyadic interactions between Hungarian speakers to see whether speakers show priming in the use of the long/short form of the indefinite subjunctive and whether their choices reflect variation in the ambient language. For each target variant of the subjunctive, I identified the last preceding subjunctive variant as its prime.</p>
<p>I tested the following research questions using a series of hierarchical Bayesian models (see the Appendix for the full model comparison). Each question maps onto a specific model comparison.</p>
<disp-quote>
<list list-type="simple">
<list-item><p><bold>RQ1</bold>. Does the use of a subjunctive variant (the <italic>target</italic>) depend on the last used variant (the <italic>prime</italic>)? I compared a model with the prime effect to a baseline model with only lexical frequency as a predictor.</p></list-item>
<list-item><p><bold>RQ2</bold>. Is this priming effect restricted to repetitions of the same verb, or does it generalise across verbs? I compared a model with a prime &#215; same-verb interaction to a model with only a main effect of prime.</p></list-item>
<list-item><p><bold>RQ3</bold>. Does priming differ depending on whether the same or a different speaker produced the prime? I compared a model with a prime &#215; speaker interaction to the best model without it.</p></list-item>
<list-item><p><bold>RQ4</bold>. Is a preference for the long/short form shaped by the verb&#8217;s lexical distribution in the webcorpus? The corpus frequency predictor is present in all models; I assess its contribution to the best model.</p></list-item>
</list>
</disp-quote>
<p>These questions are descriptive. The pattern of results can be interpreted in different theoretical frameworks. If priming is restricted to the same verb and the same inflected form, this is consistent with whole-word facilitation (lexical boost). If priming generalises across verbs sharing the same suffix, this would point to affix-level or paradigm-level convergence. I return to this in Section 3.</p>
<sec>
<title>2.1 Sources</title>
<p>I drew corpus data from a frequency list based on the second Hungarian Webcorpus (<xref ref-type="bibr" rid="B25">Nemeskey, 2020</xref>; <xref ref-type="bibr" rid="B31">R&#225;cz, 2025</xref>), which was built from Common Crawl and has a size of 9 billion words. I used transcribed conversation data from two sources. The Akaka Maptask Corpus consists of five hours of recordings of 24 task-oriented dialogues recorded using head-mounted microphones, with pairs of participants completing a map task (<xref ref-type="bibr" rid="B24">Moln&#225;r et al., 2023</xref>). The dataset has recordings from 46 participants (24 women, 22 men, mean age 20). The Budapest Games Corpus consists of 9 hours of 36 dialogues of pairs of participants. The dataset has recordings from 12 participants (five women, seven men), working together in an object identification task. Participants were recruited using convenience sampling, were matched in age group between the ages of 20 and 60 and knew each other (<xref ref-type="bibr" rid="B20">M&#225;dy et al., 2023</xref>). For more details, see Mihajlik et al. (<xref ref-type="bibr" rid="B23">2024</xref>).</p>
<p>The distributions of the raw data can be seen in <xref ref-type="fig" rid="F1">Figure 1</xref>. We see the number of long/short forms used by speakers across conversations on the left and the number of long/short forms per verb lemma in total on the right. Both show a Pareto-distribution. While many conversations see only sporadic use of the subjunctive, in some, participants use the subjunctive extensively. Some verbs, like &#8216;wait&#8217;, &#8216;go&#8217;, or &#8216;watch&#8217;, are overrepresented in subjunctive use.</p>
<fig id="F1">
<caption>
<p><bold>Figure 1:</bold> Counts of long/short indefinite subjunctive forms in the raw data. (i) Number of long/short forms used by Speaker 1/2 in each conversation, (ii) number of long/short forms per verb lemma across all conversations (excluding hapaxes).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="glossapx-5-1-50683-g1.png"/>
</fig>
</sec>
<sec>
<title>2.2 Methods</title>
<p>For details on the transcription and annotation of the original data, see M&#225;dy et al. (<xref ref-type="bibr" rid="B20">2023</xref>) and Moln&#225;r et al. (<xref ref-type="bibr" rid="B24">2023</xref>). I used force-aligned Praat textgrids and extracted the word tier and speaker identifier from each textgrid file. I filtered the resulting dataset to only include indefinite subjunctives and checked the data to avoid overlaps between subjunctive forms.</p>
<p>I drew a dataset of subjunctive forms from the webcorpus and calculated log(freq(long variant)/freq(short variant)) for each lemma.</p>
<p>I joined these data with the subjunctive dataset from the conversations to add word-level information. Three observations, for three different low-frequency verbs (<italic>tenyerel</italic> &#8216;palm&#8217; <italic>kalauzol</italic> &#8216;guide&#8217;, <italic>tologat</italic> &#8216;push habitually&#8217;) in the conversation data had no long form in the corpus. I excluded these from the final analysis.</p>
<p>The final dataset consisted of 334 subjunctives from 22 recordings in the Akaka Maptask Corpus and 27 recordings in the Budapest Games Corpus. This is 0.04% of all transcribed words in the recordings. The sample was hand-checked for accuracy.</p>
</sec>
<sec>
<title>2.3 Data analysis</title>
<p>I analysed the data in R (<xref ref-type="bibr" rid="B30">R Core Team, 2025</xref>), fit models in brms (<xref ref-type="bibr" rid="B6">B&#252;rkner, 2021</xref>), and used the ggplot2 and sjPlot packages for visualisations (<xref ref-type="bibr" rid="B19">L&#252;decke, 2025</xref>; <xref ref-type="bibr" rid="B42">Wickham, 2016</xref>).</p>
<p>The analysis compares each subjunctive verb form (the target) with the last subjunctive verb form (the prime) in the conversation. The distance between the prime and the target can vary. Out of the 334 observations, 3 are infrequent verbs that have no long form in the webcorpus and 52 are single mentions or first mentions in a conversation, meaning that there is no preceding variant to compare them with. This leaves 279 observations which belong to 26 verbs. The top five verbs (<italic>megy</italic> &#8216;go&#8217;, <italic>v&#225;r</italic> &#8216;wait&#8217;, <italic>figyel</italic> &#8216;watch&#8217;, <italic>j&#246;n</italic> &#8216;come&#8217;, and <italic>indul</italic> &#8216;start&#8217;) are responsible for 80% of all observations, with a Gini coefficient of 0.76 (see <xref ref-type="fig" rid="F1">Figure 1</xref>). The Gini coefficient quantifies distributional inequality: here, it indicates that most tokens belong to few types.</p>
<p>I fit a series of eight hierarchical generalised linear models with a binomial error distribution and a logit link function, predicting <italic>p</italic>(target is long). All models included conversation, speaker, and verb lemma as grouping factors (random intercepts). Fixed effects varied across models: the baseline model included only the verb&#8217;s corpus log-odds of the long form; subsequent models added the prime (long/short), pairwise interactions of the prime with same/different verb, same/different speaker, distance, and corpus frequency, as well as a three-way interaction and a model with all two-way interactions. The full set of models and their formulae are reported in the Appendix (Table A1).</p>
<p>I placed Normal(0, 3) priors on intercepts, Normal(0, 2) priors on fixed-effect coefficients, and Exponential(1) priors on random-effect standard deviations. These are weakly informative priors that allow large effects on the log-odds scale, while regularising against implausible extremes, following standard recommendations for logistic regression (<xref ref-type="bibr" rid="B10">Gelman et al., 2008</xref>). I compared models using approximate leave-one-out cross-validation (LOO-CV; <xref ref-type="bibr" rid="B40">Vehtari et al., 2017</xref>) and Bayes Factors computed via bridge sampling (<xref ref-type="bibr" rid="B14">Gronau et al., 2020</xref>). Residual autocorrelation was inspected manually and was not detected in any model.</p>
<p>I report the best-fitting model below. Where relevant, I convert log-odds estimates to probabilities using the inverse logit function (<italic>p</italic> = exp(<italic>x</italic>)/(1 + exp(<italic>x</italic>)), implemented as plogis() in R) to aid interpretation. I use three significant figures throughout (unless further precision is informative), matching <xref ref-type="table" rid="T1">Table 1</xref>.</p>
<table-wrap id="T1">
<caption>
<p><bold>Table 1</bold> Estimates of the best model, with 95% credible intervals.</p>
</caption>
<table>
<tbody>
<tr>
<td align="left" valign="top"><bold>term</bold></td>
<td align="left" valign="top"><bold>estimate</bold></td>
<td align="left" valign="top" colspan="2"><bold>95% CI</bold></td>
</tr>
<tr>
<td align="left" valign="top">intercept</td>
<td align="left" valign="top">&#8211;4.85</td>
<td align="left" valign="top">&#8211;7.12</td>
<td align="left" valign="top">&#8211;2.67</td>
</tr>
<tr>
<td align="left" valign="top">prime is long</td>
<td align="left" valign="top">1.68</td>
<td align="left" valign="top">0.70</td>
<td align="left" valign="top">2.66</td>
</tr>
<tr>
<td align="left" valign="top">prime is different verb</td>
<td align="left" valign="top">1.65</td>
<td align="left" valign="top">0.76</td>
<td align="left" valign="top">2.58</td>
</tr>
<tr>
<td align="left" valign="top">target log(long/short) in webcorpus</td>
<td align="left" valign="top">4.15</td>
<td align="left" valign="top">1.31</td>
<td align="left" valign="top">6.87</td>
</tr>
<tr>
<td align="left" valign="top">prime long form &#215; different verb</td>
<td align="left" valign="top">&#8211;2.00</td>
<td align="left" valign="top">&#8211;3.29</td>
<td align="left" valign="top">&#8211;0.71</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec>
<title>2.4 Results</title>
<p>The estimates of the best model can be seen in <xref ref-type="table" rid="T1">Table 1</xref>. Both priming effects and lexical effects are visible.</p>
<p>The best-fitting model (fit3 in the Appendix) includes the prime, whether the prime is the same or a different verb, their interaction, and the verb&#8217;s corpus log-odds. Model comparison supports the following answers to the research questions.</p>
<disp-quote>
<p><italic>RQ1: Priming</italic>. The model with within-verb priming is strongly favoured over the baseline (BF = 62.26; Table A1). A long prime increases the probability of a long target.</p>
<p><italic>RQ2: Same verb vs. across verbs</italic>. The interaction model is strongly favoured over a model with only a main effect of prime (BF = 106.72). Priming is restricted to the same verb lemma. When the previous verb was different, the priming effect was not credibly different from zero (95% CI: [&#8211;1.32; 0.63], calculated from the joint posterior).</p>
<p><italic>RQ3: Speaker identity</italic>. Adding a prime &#215; speaker interaction did not improve model fit (BF = 0.008 in favour of the more complex model). The data provide no evidence that priming differs depending on who produced the prime. I return to the interpretation of this null result in Section 3.</p>
<p><italic>RQ4: Lexical distribution</italic>. The verb&#8217;s corpus log-odds of the long form is a strong predictor in all models (estimate: 4.15, 95% CI: [1.31; 6.87]). Additionally, prime &#215; distance (BF = 0.002) and prime &#215; corpus frequency (BF = 0.006) did not improve fit over the best model. The priming and lexical effects are additive. A three-way interaction of prime, same verb, and corpus frequency yielded indeterminate evidence (BF = 0.82). Full model comparisons are reported in the Appendix.</p>
</disp-quote>
<p>Going back to the best model, the credible intervals capture how the groups are related to the intercept (short primes). How they are related to each other is best understood looking at the raw data in <xref ref-type="fig" rid="F2">Figure 2</xref> and the model predictions in <xref ref-type="fig" rid="F3">Figure 3</xref>.</p>
<fig id="F2">
<caption>
<p><bold>Figure 2:</bold> Verb long/short preference in the raw data. (i) Targets are more likely to be long in conversations (y axis) if the verb is more likely to be long in the webcorpus (x axis &#8211; the widening of the credible interval reflects gaps in the distribution of verb types along the corpus frequency axis). (ii) Long primes increase the probability of long targets; in the absence of a long prime, the short form predominates. This effect is absent if it was a different lemma.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="glossapx-5-1-50683-g2.png"/>
</fig>
<fig id="F3">
<caption>
<p><bold>Figure 3:</bold> Verb long/short preference based on best model. (i) Targets are more likely to be long in conversations (y axis) if the verb is more likely to be long in the webcorpus (x axis), (ii) Long primes increase the probability of long targets if the prime was the same verb lemma. In the absence of a long prime, short forms predominate. This effect is absent if it was a different lemma.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="glossapx-5-1-50683-g3.png"/>
</fig>
<p>If a verb prefers the long form in the webcorpus (x axis), it will do so in the conversations (y axis). We see a strong priming effect within the same verb: a long prime is followed by a long target. In the absence of a long prime, the short form predominates. There is no clear priming pattern across verbs.</p>
<p>Both the lexical effect and the priming effect persist in the best model. If we transform log odds to probabilities using the inverse logit function, we can say that verbs at the highest end of the (scaled) corpus distribution of long forms are about 30x more likely to pick a long form over verbs at the lowest end (<italic>plogis</italic>(&#8211;4.85 + 4.15) = .332 over <italic>plogis</italic>(&#8211;4.85) = .008). Within verbs, a long prime variant means that a long target variant is about 3x more likely (<italic>plogis</italic>(&#8211;4.85 + 1.68) = .040 over <italic>plogis</italic>(&#8211;4.85) = .008). The two effects are additive, not interactive, in the model. Model predictions can be seen in <xref ref-type="fig" rid="F3">Figure 3</xref>.</p>
<p>Based on model selection, the best-fitting model has two plausible alternatives. First, target verb length may reflect lexical preferences alone, with no priming effect present. There is robust evidence against this model (BF = 62.61).</p>
<p>A second alternative model includes an interaction between prime and lexical preference. This is, again, best understood by looking at the predictions of this alternative model (<xref ref-type="fig" rid="F4">Figure 4</xref>). In this model, it remains true that the prime only has a clear effect on the target if it is the same verb. However, this effect interacts with the verb&#8217;s lexical preference: if a verb is unlikely to be long based on the webcorpus, it will remain short in the conversation, irrespective of the prime. If the verb is more likely to be long in the corpus, this is more prominent in the conversation. Evidence for this interaction is indeterminate (BF = 0.82 relative to the simpler model), meaning the data neither support nor rule out a role for lexical expectations in modulating priming. If this interaction proved robust in larger samples, it would be consistent with expectation-based models of production. I return to this in Section 3.</p>
<fig id="F4">
<caption>
<p><bold>Figure 4:</bold> Predictions of alternate model: interaction between lexical preference, prime long/short, and verb lemma. If the lemma is the same (left), we see a stronger lexical effect for long primes versus short primes. If the lemma is different, this effect is absent (right).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="glossapx-5-1-50683-g4.png"/>
</fig>
<p>Leave-one-out cross-validation fails to discriminate between models (identical LOO scores across all three candidates), likely reflecting data scarcity: with relatively few observations and few verb types, out-of-sample prediction is uninformative. As noted above, Bayes Factor comparison favours the priming model over the null, but remains indeterminate between the priming model and the more complex alternative which assumes a lexical preference : priming interaction. This indeterminacy may also reflect insufficient data to justify additional parameters and could be tested on a larger interactional dataset.</p>
<p>Since it was flagged by cross-validation, I used a permutation test to further confirm the priming effect (the reported model). I subsetted the data to observations where the target was the same verb as the prime (138 observations). I randomly shuffled the prime (long/short) 10,000 times and calculated a simulated ratio for the number of long targets over short targets. The mean of these ratios was 0.49. The true mean of long targets following long primes was much higher: 0.78. No simulated mean was higher than this. This means that the priming effect across identical primes and targets is unlikely to be due to chance alone (<italic>p</italic> &#60; 0.001). This simple test does not replicate the entire interaction in the best model, but it does lend further support to the robustness of the priming effect.</p>
</sec>
</sec>
<sec>
<title>3. Conclusion</title>
<p>I analysed spoken interactions in Hungarian-speaking dyads working together in simple game-like tasks, looking at variation in the 2<sc>sg</sc> indefinite subjunctive, which is largely used to express requests and commands in Hungarian (<xref ref-type="bibr" rid="B39">T&#243;th, 2007</xref>). I wanted to see whether the choice of the short or long variant of subjunctive forms persisted across spoken dialogue.</p>
<p>I found that a long prime variant made the use of a long target variant about 3x more likely for the same verb. In the absence of a long prime, the short form predominated. The effect was restricted to repetitions of the same verb lemma and did not generalise across verbs. In addition, the choice of the short or long subjunctive for a given verb strongly correlated with its general preference for short or long subjunctives in a large webcorpus of Hungarian. The priming and lexical effects were additive.</p>
<p>The speaker distinction did not improve model fit (BF = 0.008). This means that the data do not distinguish between reciprocal priming, in which one speaker&#8217;s output facilitates the other&#8217;s production, and parallel self-priming, in which both speakers are independently influenced by the same lexical distributions. Under the strong frequency effects observed here, both speakers are likely to gravitate toward the same form, regardless of interpersonal coordination. Distinguishing these accounts would require designs that manipulate speaker exposure independently, such as confederate paradigms or asymmetric information tasks.</p>
<p>The finding that priming operates exclusively within lemmata is consistent with a lexical boost account, in which facilitation is driven by recent activation of the same word form (<xref ref-type="bibr" rid="B27">Pickering &#38; Branigan, 1998</xref>). The priming observed here operates on specific inflected tokens: the same verb in the same subjunctive variant. There is no evidence of generalisation across verbs, which would be expected under affix-level or paradigm-level convergence. We therefore cannot conclude that the effect operates at the level of subword representations. It may equally reflect whole-word storage and retrieval.</p>
<p>At the same time, the alternation under study is morphologically conditioned: the long and short forms are distinct inflectional variants of the same exponent. The strong corpus-frequency effect shows that distributional properties of the inflectional paradigm shape conversational choices independently of local priming. This is a necessary, though not sufficient, condition for morphological-level processes. A stronger test would require evidence that priming of the long form of one verb increases the probability of the long form for a different verb with a similar distributional profile. The data do not show this, but the sample is small and heavily concentrated in a few high-frequency verbs.</p>
<p>We discussed an alternate, more complex model of the results, in which lexical preference and prime jointly shape target choice in conversation. Evidence for this interaction is indeterminate in the data (BF = 0.82). If the interaction between lexical preference and prime length proves robust in larger samples, it would match the broader consensus regarding the relationship between priming effects and lexical structure in morphological processes (<xref ref-type="bibr" rid="B15">Hay &#38; Baayen, 2005</xref>) and support expectation-based models of production in which speakers adjust form selection to adapt to the statistics of the current environment (<xref ref-type="bibr" rid="B17">Jaeger &#38; Snider, 2013</xref>). Such models track expectation violations and converge towards higher-probability forms. However, evidence for these mechanisms is contested (<xref ref-type="bibr" rid="B7">Fazekas et al., 2024</xref>), and the conclusions of the data are indeterminate on this point, potentially due to sample size limitations.</p>
<p>A caveat on the interpretation of additive effects in logistic regression: because the logit link function is nonlinear, effects that are additive on the log-odds scale are not additive on the probability scale (<xref ref-type="bibr" rid="B1">Ai &#38; Norton, 2003</xref>). In the model, the prime effect and the corpus-frequency effect do not interact on the log-odds scale. On the probability scale, however, the same prime coefficient produces a larger shift in probability for verbs with a higher baseline probability of the long form. This means that the prime&#8217;s influence on variant choice is already modulated by lexical frequency, even without an explicit interaction term. This does not amount to a surprisal-based mechanism, but it does mean that the &#8220;additive&#8221; result is less theoretically flat than it may appear: the model already captures a form of frequency-dependent priming through the geometry of the link function. This may also explain why the explicit interaction model (fit7) is indeterminate, rather than clearly disfavoured: it is trying to capture variance that the logit link already partially accounts for.</p>
<p>Large-scale online morphological convergence tasks reported in R&#225;cz and Luk&#225;cs (<xref ref-type="bibr" rid="B33">2024</xref>) and R&#225;cz et al. (<xref ref-type="bibr" rid="B32">2020</xref>) find this interaction between lexical distribution and linguistic convergence. In those tasks, a nonword&#8217;s behaviour is shaped by the interaction of the nonword&#8217;s baseline distribution and the behaviour of the co-player. At the same time, these online tasks expose a larger number of participants to a much wider, well-balanced array of nonwords, making it much easier to identify relevant complex patterns in the data. It is entirely possible that a larger conversational dataset would also support the otherwise intuitive result that priming effects and lexical effects interact.</p>
<p>The main strength of the analysis is that it demonstrates priming of morphological variants for existing words within a realistic language interaction. Work on morphological variation and priming in naturalistic settings is scarce in the literature. This study provides important confirmation for existing corpus-based (<xref ref-type="bibr" rid="B38">Szmrecsanyi, 2006</xref>; <xref ref-type="bibr" rid="B41">Weiner &#38; Labov, 1983</xref>) and artificial language studies (<xref ref-type="bibr" rid="B33">R&#225;cz &#38; Luk&#225;cs, 2024</xref>; <xref ref-type="bibr" rid="B32">R&#225;cz et al., 2020</xref>) on morphological convergence, and extends these findings to a morphologically rich language.</p>
<p>The study contends with a number of limitations. The interactions play out between familiar participants in an informal setting. Interactive tasks also carry a lot of emphasis, with participants repeatedly telling one another to look at something or go somewhere, and emphasis might explain some repetition in the long/short form of the subjunctive in these conversations. The stylistic and social factors driving variation are not known. The relatively small set of verbs and the structure of the conversations likely translates to more limited insights into the relationship between priming effects and lexical effects on morphological convergence. This is not unusual, as state-of-the-art theories assume robust but hard-to-detect pressures acting on language variation and change (<xref ref-type="bibr" rid="B16">Hay et al., 2015</xref>). The results remain important as a demonstration of inflected-form priming in a naturalistic interaction, with lexical distributions independently scaffolding variant choice.</p>
<p>Any quantitative analysis of convergence effects based on audio transcripts carries more general limitations, and these reference two important questions in current theory.</p>
<p>First, speaker convergence across variable linguistic output is related to more general work on variation and systematising in language and cognition. Schumacher and Pierrehumbert (<xref ref-type="bibr" rid="B35">2021</xref>) make the point that linguistic behaviour that looks like probability matching in the aggregate might stem from pooling heterogeneous speakers, who modulate their choices based on socially learned baselines and associations. The aggregate analysis of the linguistic data, despite hierarchical modelling, might pave over a similar set of individual strategies that determine morphological convergence above and beyond repetition and lexical pressure. Schumacher and Pierrehumbert also make the observation that formulating a rule might be a separate, subsequent, step after learning from example, and this fits in with the result that priming and lexical effects are additive, rather than interactive, in the dataset. It is possible that participants do use rules in selecting the subjunctive for various verbs, but these rules are not directly built on corpus frequencies. This would imply that there might be an interaction between lexical preference and priming in the dataset, but this is poorly represented by aggregate frequency counts drawn from a webcorpus.</p>
<p>Second, there is an emerging consensus in linguistics that the efficient communication of information moulds language variation down to the micro-level (<xref ref-type="bibr" rid="B26">Piantadosi et al., 2011</xref>). Analysing dyadic interactions should be a stark reminder that much of this information is non-linguistic and hard to operationalise or quantify. Joint gaze, attention, non-verbal turn-taking, and the larger social context of any interaction could have as much influence on morphological convergence as lexical distributions or word-to-word priming. This highlights the lessons that quantitative psycholinguistics can learn from linguistic anthropology and ethnographic fieldwork, which can capture the nonverbal and social dynamics that shape human interactions (<xref ref-type="bibr" rid="B12">Goodwin, 2018</xref>; <xref ref-type="bibr" rid="B37">Stivers et al., 2009</xref>).</p>
</sec>
</body>
<back>
<sec>
<title>Appendix</title>
<table-wrap id="T2">
<caption>
<p><bold>Table A1:</bold> Models fit on the data.</p>
</caption>
<table>
<tbody>
<tr>
<td align="left" valign="top"><bold>Description</bold></td>
<td align="left" valign="top"><bold>Formula</bold></td>
<td align="left" valign="top"><bold>Fit</bold></td>
<td align="left" valign="top"><bold>elpd_diff</bold></td>
<td align="left" valign="top"><bold>se_diff</bold></td>
<td align="left" valign="top"><bold>bfs</bold></td>
<td align="left" valign="top"><bold>bf_labels</bold></td>
</tr>
<tr>
<td align="left" valign="top">Baseline (lexical)</td>
<td align="left" valign="top">&#8764; corpus log(long/short)</td>
<td align="left" valign="top">fit1</td>
<td align="left" valign="top">&#8211;2.35</td>
<td align="left" valign="top">4.00</td>
<td align="left" valign="top">62.26</td>
<td align="left" valign="top">fit3/fit1</td>
</tr>
<tr>
<td align="left" valign="top">Prime main effect</td>
<td align="left" valign="top">&#8764; long prime + corpus log(long/short)</td>
<td align="left" valign="top">fit2</td>
<td align="left" valign="top">&#8211;4.62</td>
<td align="left" valign="top">3.57</td>
<td align="left" valign="top">106.72</td>
<td align="left" valign="top">fit3/fit2</td>
</tr>
<tr>
<td align="left" valign="top">Prime &#215; same verb</td>
<td align="left" valign="top">&#8764; long prime * prime diff. lemma + corpus log(long/short)</td>
<td align="left" valign="top">fit3</td>
<td align="left" valign="top">0.00</td>
<td align="left" valign="top">0.00</td>
<td align="left" valign="top">&#8211;</td>
<td align="left" valign="top">&#8211;</td>
</tr>
<tr>
<td align="left" valign="top">Prime &#215; speaker</td>
<td align="left" valign="top">&#8764; long prime * prime diff. speaker + corpus log(long/short) + prime diff. lemma</td>
<td align="left" valign="top">fit4</td>
<td align="left" valign="top">&#8211;3.02</td>
<td align="left" valign="top">2.73</td>
<td align="left" valign="top">0.008</td>
<td align="left" valign="top">fit4/fit3</td>
</tr>
<tr>
<td align="left" valign="top">Prime &#215; distance</td>
<td align="left" valign="top">&#8764; long prime * distance to prime + corpus log(long/short) + prime diff. lemma + prime diff. speaker</td>
<td align="left" valign="top">fit5</td>
<td align="left" valign="top">&#8211;5.13</td>
<td align="left" valign="top">3.15</td>
<td align="left" valign="top">0.002</td>
<td align="left" valign="top">fit5/fit3</td>
</tr>
<tr>
<td align="left" valign="top">Prime &#215; frequency</td>
<td align="left" valign="top">&#8764; long prime * corpus log(long/short) + prime diff. lemma + prime diff. speaker</td>
<td align="left" valign="top">fit6</td>
<td align="left" valign="top">&#8211;4.83</td>
<td align="left" valign="top">3.08</td>
<td align="left" valign="top">0.006</td>
<td align="left" valign="top">fit6/fit3</td>
</tr>
<tr>
<td align="left" valign="top">Three-way interaction</td>
<td align="left" valign="top">&#8764; long prime * prime diff. lemma * corpus log(long/short) + prime diff. speaker</td>
<td align="left" valign="top">fit7</td>
<td align="left" valign="top">&#8211;0.15</td>
<td align="left" valign="top">1.20</td>
<td align="left" valign="top">0.82</td>
<td align="left" valign="top">fit7/fit3</td>
</tr>
<tr>
<td align="left" valign="top">All two-way int.</td>
<td align="left" valign="top">&#8764; long prime * prime diff. lemma + long prime * corpus log(long/short) + long prime * prime diff. speaker + long prime * distance to prime</td>
<td align="left" valign="top">fit8</td>
<td align="left" valign="top">&#8211;2.39</td>
<td align="left" valign="top">1.15</td>
<td align="left" valign="top">0.05</td>
<td align="left" valign="top">fit8/fit3</td>
</tr>
</tbody>
</table>
</table-wrap>
<p><xref ref-type="table" rid="T2">Table A1</xref> summarises the eight hierarchical generalised linear models fit to the data. All models predict p(target is long) with a binomial error distribution and a logit link. All models include conversation, speaker, and verb lemma as grouping factors (random intercepts). Priors are weakly informative: Normal(0, 3) on the intercept, Normal(0, 2) on fixed-effect coefficients, and Exponential(1) on random-effect standard deviations. These follow standard recommendations for logistic regression (<xref ref-type="bibr" rid="B10">Gelman et al., 2008</xref>) and are intended to regularise estimates without imposing strong directional constraints.</p>
<p>Models were compared using approximate leave-one-out cross-validation (LOO-CV; <xref ref-type="bibr" rid="B40">Vehtari et al., 2017</xref>) and Bayes Factors (BF), computed via bridge sampling as implemented in the bayestestR package (<xref ref-type="bibr" rid="B21">Makowski et al., 2019</xref>). The elpd_diff column reports the difference in expected log pointwise predictive density, relative to the best-fitting model (fit3); negative values indicate worse out-of-sample prediction. The se_diff column gives the standard error of this difference. LOO-CV differences are small, relative to their standard errors across all models, which reflects limited discriminability, given sample size (n = 279).</p>
<p>Bayes Factors are reported as the ratio of marginal likelihoods for the row model over its comparison model (bf_labels column). Fit3, the prime &#215; same-verb model, serves as the reference. BF &#62; 10 indicates strong evidence in favour of the numerator model; BF &#60; 0.1 indicates strong evidence against it; values near 1 are indeterminate. Fit3 is strongly favoured over the baseline (BF = 62.26) and over the prime-only model (BF = 106.72), confirming that within-verb priming improves on both lexical effects alone and an undifferentiated prime effect. Adding speaker identity (fit4, BF = 0.008), prime-target distance (fit5, BF = 0.002), prime &#215; corpus frequency (fit6, BF = 0.006), or all two-way interactions (fit8, BF = 0.05) does not improve on fit3 and is actively disfavoured. The three-way interaction model (fit7, BF = 0.82) is indeterminate: the data neither support nor rule out an additional interaction between priming, lemma identity, and corpus frequency.</p>
<p>Residual autocorrelation was checked for all models and was not detected.</p>
</sec>
<sec>
<title>Abbreviations</title>
<table-wrap>
<table>
<tbody>
<tr>
<td align="left" valign="top"><bold>2SG</bold></td>
<td align="left" valign="top">second person singular</td>
</tr>
<tr>
<td align="left" valign="top"><bold>3SG</bold></td>
<td align="left" valign="top">third person singular</td>
</tr>
<tr>
<td align="left" valign="top"><bold>ACC</bold></td>
<td align="left" valign="top">accusative</td>
</tr>
<tr>
<td align="left" valign="top"><bold>INDEF</bold></td>
<td align="left" valign="top">indefinite</td>
</tr>
<tr>
<td align="left" valign="top"><bold>PART</bold></td>
<td align="left" valign="top">particle</td>
</tr>
<tr>
<td align="left" valign="top"><bold>SBJV</bold></td>
<td align="left" valign="top">subjunctive</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec>
<title>Data accessibility statement</title>
<p>Data and code are available at <ext-link xmlns:xlink="http://www.w3.org/1999/xlink" ext-link-type="uri" xlink:href="https://doi.org/10.5281/zenodo.20204995">https://doi.org/10.5281/zenodo.20204995</ext-link>.</p>
</sec>
<sec>
<title>Ethics and consent</title>
<p>This analysis uses secondary data sources. Data made available for the replication of this analysis is fully anonymised.</p>
</sec>
<sec>
<title>Funding</title>
<p>Bolyai J&#225;nos Research Scholarship of the Hungarian Academy of Sciences.</p>
</sec>
<sec>
<title>Acknowledgements</title>
<p>I would like to thank my Editor, my Reviewers, Tekla Gr&#225;czi, and M&#225;rton S&#243;skuthy.</p>
</sec>
<sec>
<title>Competing interests</title>
<p>The author has no competing interests to declare.</p>
</sec>
<sec>
<title>ORCiD IDs</title>
<p><bold>P&#233;ter R&#225;cz:</bold> <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://orcid.org/0000-0001-7896-4801">https://orcid.org/0000-0001-7896-4801</ext-link></p>
</sec>
<ref-list>
<ref id="B1"><mixed-citation publication-type="journal"><string-name><surname>Ai</surname>, <given-names>C.</given-names></string-name>, &#38; <string-name><surname>Norton</surname>, <given-names>E. C.</given-names></string-name> (<year>2003</year>). <article-title>Interaction terms in logit and probit models</article-title>. <source>Economics Letters</source>, <volume>80</volume> (<issue>1</issue>), <fpage>123</fpage>&#8211;<lpage>129</lpage>. <pub-id pub-id-type="doi">10.1016/S0165-1765(03)00032-6</pub-id></mixed-citation></ref>
<ref id="B2"><mixed-citation publication-type="journal"><string-name><surname>Amenta</surname>, <given-names>S.</given-names></string-name>, &#38; <string-name><surname>Crepaldi</surname>, <given-names>D.</given-names></string-name> (<year>2012</year>). <article-title>Morphological processing as we know it: An analytical review of morphological effects in visual word identification</article-title>. <source>Frontiers in Psychology</source>, <volume>3</volume>, <elocation-id>232</elocation-id>. <pub-id pub-id-type="doi">10.3389/fpsyg.2012.00232</pub-id></mixed-citation></ref>
<ref id="B3"><mixed-citation publication-type="journal"><string-name><surname>Babel</surname>, <given-names>M.</given-names></string-name> (<year>2012</year>). <article-title>Evidence for phonetic and social selectivity in spontaneous phonetic imitation</article-title>. <source>Journal of Phonetics</source>, <volume>40</volume> (<issue>1</issue>), <fpage>177</fpage>&#8211;<lpage>189</lpage>. <pub-id pub-id-type="doi">10.1016/j.wocn.2011.09.001</pub-id></mixed-citation></ref>
<ref id="B4"><mixed-citation publication-type="journal"><string-name><surname>Beckner</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Blythe</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Bybee</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Christiansen</surname>, <given-names>M. H.</given-names></string-name>, <string-name><surname>Croft</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Ellis</surname>, <given-names>N. C.</given-names></string-name>, <string-name><surname>Holland</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Ke</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Larsen-Freeman</surname>, <given-names>D.</given-names></string-name>, et al. (<year>2009</year>). <article-title>Language is a complex adaptive system: Position paper</article-title>. <source>Language Learning</source>, <volume>59</volume>, <fpage>1</fpage>&#8211;<lpage>26</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9922.2009.00533.x</pub-id></mixed-citation></ref>
<ref id="B5"><mixed-citation publication-type="journal"><string-name><surname>Bock</surname>, <given-names>J. K.</given-names></string-name> (<year>1986</year>). <article-title>Syntactic persistence in language production</article-title>. <source>Cognitive Psychology</source>, <volume>18</volume> (<issue>3</issue>), <fpage>355</fpage>&#8211;<lpage>387</lpage>. <pub-id pub-id-type="doi">10.1016/0010-0285(86)90004-6</pub-id></mixed-citation></ref>
<ref id="B6"><mixed-citation publication-type="journal"><string-name><surname>B&#252;rkner</surname>, <given-names>P.-C.</given-names></string-name> (<year>2021</year>). <article-title>Bayesian item response modeling in R with brms and Stan</article-title>. <source>Journal of Statistical Software</source>, <volume>100</volume> (<issue>5</issue>), <fpage>1</fpage>&#8211;<lpage>54</lpage>. <pub-id pub-id-type="doi">10.18637/jss.v100.i05</pub-id></mixed-citation></ref>
<ref id="B7"><mixed-citation publication-type="journal"><string-name><surname>Fazekas</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Sala</surname>, <given-names>G.</given-names></string-name>, &#38; <string-name><surname>Pine</surname>, <given-names>J.</given-names></string-name> (<year>2024</year>). <article-title>Prime surprisal as a tool for assessing error-based learning theories: A systematic review</article-title>. <source>Languages</source>, <volume>9</volume> (<issue>4</issue>), <elocation-id>147</elocation-id>. <pub-id pub-id-type="doi">10.3390/languages9040147</pub-id></mixed-citation></ref>
<ref id="B8"><mixed-citation publication-type="journal"><string-name><surname>Garrod</surname>, <given-names>S.</given-names></string-name>, &#38; <string-name><surname>Pickering</surname>, <given-names>M. J.</given-names></string-name> (<year>2004</year>). <article-title>Why is conversation so easy?</article-title> <source>Trends in Cognitive Sciences</source>, <volume>8</volume> (<issue>1</issue>), <fpage>8</fpage>&#8211;<lpage>11</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2003.10.016</pub-id></mixed-citation></ref>
<ref id="B9"><mixed-citation publication-type="journal"><string-name><surname>Garrod</surname>, <given-names>S.</given-names></string-name>, &#38; <string-name><surname>Pickering</surname>, <given-names>M. J.</given-names></string-name> (<year>2009</year>). <article-title>Joint action, interactive alignment, and dialog</article-title>. <source>Topics in Cognitive Science</source>, <volume>1</volume> (<issue>2</issue>), <fpage>292</fpage>&#8211;<lpage>304</lpage>. <pub-id pub-id-type="doi">10.1111/j.1756-8765.2009.01020.x</pub-id></mixed-citation></ref>
<ref id="B10"><mixed-citation publication-type="journal"><string-name><surname>Gelman</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Jakulin</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Pittau</surname>, <given-names>M. G.</given-names></string-name>, &#38; <string-name><surname>Su</surname>, <given-names>Y.-S.</given-names></string-name> (<year>2008</year>). <article-title>A weakly informative default prior distribution for logistic and other regression models</article-title>. <source>The Annals of Applied Statistics</source>, <volume>2</volume> (<issue>4</issue>), <fpage>1360</fpage>&#8211;<lpage>1383</lpage>. <pub-id pub-id-type="doi">10.1214/08-AOAS191</pub-id></mixed-citation></ref>
<ref id="B11"><mixed-citation publication-type="book"><string-name><surname>Giles</surname>, <given-names>H.</given-names></string-name>, &#38; <string-name><surname>Coupland</surname>, <given-names>N.</given-names></string-name> (<year>1991</year>). <source>Language: Contexts and consequences</source>. <publisher-name>Thomson Brooks/Cole Publishing Co.</publisher-name></mixed-citation></ref>
<ref id="B12"><mixed-citation publication-type="book"><string-name><surname>Goodwin</surname>, <given-names>C.</given-names></string-name> (<year>2018</year>). <source>Co-operative action</source>. <publisher-name>Cambridge University Press</publisher-name>. <pub-id pub-id-type="doi">10.1017/9781139016735</pub-id></mixed-citation></ref>
<ref id="B13"><mixed-citation publication-type="journal"><string-name><surname>Gries</surname>, <given-names>S. T.</given-names></string-name> (<year>2005</year>). <article-title>Syntactic priming: A corpus-based approach</article-title>. <source>Journal of Psycholinguistic Research</source>, <volume>34</volume> (<issue>4</issue>), <fpage>365</fpage>&#8211;<lpage>399</lpage>. <pub-id pub-id-type="doi">10.1007/s10936-005-6139-3</pub-id></mixed-citation></ref>
<ref id="B14"><mixed-citation publication-type="journal"><string-name><surname>Gronau</surname>, <given-names>Q. F.</given-names></string-name>, <string-name><surname>Singmann</surname>, <given-names>H.</given-names></string-name>, &#38; <string-name><surname>Wagenmakers</surname>, <given-names>E.-J.</given-names></string-name> (<year>2020</year>). <article-title>Bridgesampling: An R package for estimating normalizing constants</article-title>. <source>Journal of Statistical Software</source>, <volume>92</volume> (<issue>10</issue>), <fpage>1</fpage>&#8211;<lpage>29</lpage>. <pub-id pub-id-type="doi">10.18637/jss.v092.i10</pub-id></mixed-citation></ref>
<ref id="B15"><mixed-citation publication-type="journal"><string-name><surname>Hay</surname>, <given-names>J. B.</given-names></string-name>, &#38; <string-name><surname>Baayen</surname>, <given-names>R. H.</given-names></string-name> (<year>2005</year>). <article-title>Shifting paradigms: Gradient structure in morphology</article-title>. <source>Trends in Cognitive Sciences</source>, <volume>9</volume> (<issue>7</issue>), <fpage>342</fpage>&#8211;<lpage>348</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2005.04.002</pub-id></mixed-citation></ref>
<ref id="B16"><mixed-citation publication-type="journal"><string-name><surname>Hay</surname>, <given-names>J. B.</given-names></string-name>, <string-name><surname>Pierrehumbert</surname>, <given-names>J. B.</given-names></string-name>, <string-name><surname>Walker</surname>, <given-names>A. J.</given-names></string-name>, &#38; <string-name><surname>LaShell</surname>, <given-names>P.</given-names></string-name> (<year>2015</year>). <article-title>Tracking word frequency effects through 130 years of sound change</article-title>. <source>Cognition, 139</source>, <fpage>83</fpage>&#8211;<lpage>91</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2015.02.012</pub-id></mixed-citation></ref>
<ref id="B17"><mixed-citation publication-type="journal"><string-name><surname>Jaeger</surname>, <given-names>T. F.</given-names></string-name>, &#38; <string-name><surname>Snider</surname>, <given-names>N. E.</given-names></string-name> (<year>2013</year>). <article-title>Alignment as a consequence of expectation adaptation: Syntactic priming is affected by the prime&#8217;s prediction error given both prior and recent experience</article-title>. <source>Cognition</source>, <volume>127</volume> (<issue>1</issue>), <fpage>57</fpage>&#8211;<lpage>83</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2012.10.013</pub-id></mixed-citation></ref>
<ref id="B18"><mixed-citation publication-type="book"><string-name><surname>L&#225;szl&#243;</surname>, <given-names>H.</given-names></string-name> (<year>2001</year>). <chapter-title>A felsz&#243;l&#237;t&#243; m&#243;d&#250; igealakok kett&#337;ss&#233;g&#233;nek t&#246;rt&#233;net&#233;hez [Regarding the history of the dual pattern of imperative verb forms]</chapter-title>. In <string-name><given-names>L.</given-names> <surname>B&#250;ky</surname></string-name> &#38; <string-name><given-names>T.</given-names> <surname>Forg&#225;cs</surname></string-name> (Eds.), <source>A nyelvt&#246;rt&#233;neti kutat&#225;sok &#250;jabb eredm&#233;nyei ii. magyar &#233;s finnugor alaktan [the latest results in diachronic linguistics ii: Finno-ugric morphology]</source> (pp. <fpage>55</fpage>&#8211;<lpage>65</lpage>). <publisher-name>Szegedi Tudom&#225;nyegyetem Magyar Nyelv&#233;szeti Tansz&#233;k</publisher-name>.</mixed-citation></ref>
<ref id="B19"><mixed-citation publication-type="webpage"><string-name><surname>L&#252;decke</surname>, <given-names>D.</given-names></string-name> (<year>2025</year>). <source>Sjplot: Data visualization for statistics in social science</source> [R package version <volume>2</volume>.<issue>9</issue>.<fpage>0</fpage><lpage>]</lpage>. <uri>https://CRAN.R-project.org/package=sjPlot</uri></mixed-citation></ref>
<ref id="B20"><mixed-citation publication-type="journal"><string-name><surname>M&#225;dy</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Koh&#225;ri</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Reichel</surname>, <given-names>U. D.</given-names></string-name>, <string-name><surname>Szalontai</surname>, <given-names>&#193;.</given-names></string-name>, &#38; <string-name><surname>Mihajlik</surname>, <given-names>P.</given-names></string-name> (<year>2023</year>). <article-title>The Budapest Games Corpus</article-title>. <source>Speech Research Conference</source>, <elocation-id>75</elocation-id>. <pub-id pub-id-type="doi">10.18135/BeszKutKonf.2023</pub-id></mixed-citation></ref>
<ref id="B21"><mixed-citation publication-type="journal"><string-name><surname>Makowski</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Ben-Shachar</surname>, <given-names>M. S.</given-names></string-name>, &#38; <string-name><surname>L&#252;decke</surname>, <given-names>D.</given-names></string-name> (<year>2019</year>). <article-title>Bayestestr: Describing effects and their uncertainty, existence and significance within the Bayesian framework</article-title>. <source>Journal of Open Source Software</source>, <volume>4</volume> (<issue>40</issue>), <elocation-id>1541</elocation-id>. <pub-id pub-id-type="doi">10.21105/joss.01541</pub-id></mixed-citation></ref>
<ref id="B22"><mixed-citation publication-type="book"><string-name><surname>Marslen-Wilson</surname>, <given-names>W. D.</given-names></string-name> (<year>2007</year>). <chapter-title>Morphological processes in language comprehension</chapter-title>. In <string-name><given-names>M. Gareth</given-names> <surname>Gaskell</surname></string-name> (Ed.), <source>The oxford handbook of psycholinguistics</source> (pp. <fpage>175</fpage>&#8211;<lpage>193</lpage>). <publisher-name>Oxford University Press</publisher-name>. <pub-id pub-id-type="doi">10.1093/oxfordhb/9780198568971.013.0011</pub-id></mixed-citation></ref>
<ref id="B23"><mixed-citation publication-type="journal"><string-name><surname>Mihajlik</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>M&#225;dy</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Koh&#225;ri</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Fruzsina</surname>, <given-names>F. S.</given-names></string-name>, <string-name><surname>Kiss</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Gr&#225;czi</surname>, <given-names>T. E.</given-names></string-name>, &#38; <string-name><surname>Do&#287;ru&#246;z</surname>, <given-names>A. S.</given-names></string-name> (<year>2024</year>). <article-title>Is spoken Hungarian low-resource?: A quantitative survey of Hungarian speech data sets</article-title>. <source>Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024)</source>, <fpage>9382</fpage>&#8211;<lpage>9388</lpage>. <pub-id pub-id-type="doi">10.63317/4qrohu7pa37j</pub-id></mixed-citation></ref>
<ref id="B24"><mixed-citation publication-type="journal"><string-name><surname>Moln&#225;r</surname>, <given-names>C. S.</given-names></string-name>, <string-name><surname>M&#225;dy</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Mihajlik</surname>, <given-names>P.</given-names></string-name>, &#38; <string-name><surname>Gyuris</surname>, <given-names>B.</given-names></string-name> (<year>2023</year>). <article-title>The Akaka Maptask Corpus</article-title>. <source>Besz&#233;dkutat&#225;s-Speech Research Conference</source>, <fpage>81</fpage>&#8211;<lpage>83</lpage>. <pub-id pub-id-type="doi">10.18135/BeszKutKonf.2023</pub-id></mixed-citation></ref>
<ref id="B25"><mixed-citation publication-type="thesis"><string-name><surname>Nemeskey</surname>, <given-names>D. M.</given-names></string-name> (<year>2020</year>). <source>Natural language processing methods for language modeling</source> [Doctoral dissertation]. <publisher-name>E&#246;tv&#246;s Lor&#225;nd University [ELTE Digital Institutional Repository]</publisher-name>. <pub-id pub-id-type="doi">10.15476/ELTE.2020.066</pub-id></mixed-citation></ref>
<ref id="B26"><mixed-citation publication-type="journal"><string-name><surname>Piantadosi</surname>, <given-names>S. T.</given-names></string-name>, <string-name><surname>Tily</surname>, <given-names>H.</given-names></string-name>, &#38; <string-name><surname>Gibson</surname>, <given-names>E.</given-names></string-name> (<year>2011</year>). <article-title>Word lengths are optimized for efficient communication</article-title>. <source>Proceedings of the National Academy of Sciences</source>, <volume>108</volume> (<issue>9</issue>), <fpage>3526</fpage>&#8211;<lpage>3529</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1012551108</pub-id></mixed-citation></ref>
<ref id="B27"><mixed-citation publication-type="journal"><string-name><surname>Pickering</surname>, <given-names>M. J.</given-names></string-name>, &#38; <string-name><surname>Branigan</surname>, <given-names>H. P.</given-names></string-name> (<year>1998</year>). <article-title>The representation of verbs: Evidence from syntactic priming in language production</article-title>. <source>Journal of Memory and Language</source>, <volume>39</volume> (<issue>4</issue>), <fpage>633</fpage>&#8211;<lpage>651</lpage>. <pub-id pub-id-type="doi">10.1006/jmla.1998.2592</pub-id></mixed-citation></ref>
<ref id="B28"><mixed-citation publication-type="journal"><string-name><surname>Pickering</surname>, <given-names>M. J.</given-names></string-name>, &#38; <string-name><surname>Garrod</surname>, <given-names>S.</given-names></string-name> (<year>2004</year>). <article-title>Toward a mechanistic psychology of dialogue</article-title>. <source>Behavioral and Brain Sciences</source>, <volume>27</volume> (<issue>2</issue>), <fpage>169</fpage>&#8211;<lpage>190</lpage>. <pub-id pub-id-type="doi">10.1017/S0140525X04000056</pub-id></mixed-citation></ref>
<ref id="B29"><mixed-citation publication-type="journal"><string-name><surname>Pickering</surname>, <given-names>M. J.</given-names></string-name>, &#38; <string-name><surname>Garrod</surname>, <given-names>S.</given-names></string-name> (<year>2006</year>). <article-title>Alignment as the basis for successful communication</article-title>. <source>Research on Language and Computation</source>, <volume>4</volume>(<issue>2&#8211;3</issue>), <fpage>203</fpage>&#8211;<lpage>228</lpage>. <pub-id pub-id-type="doi">10.1007/s11168-006-9004-0</pub-id></mixed-citation></ref>
<ref id="B30"><mixed-citation publication-type="webpage"><collab>R Core Team</collab>. (<year>2025</year>). <source>R: A language and environment for statistical computing</source>. R Foundation for Statistical Computing. <uri>https://www.R-project.org/</uri></mixed-citation></ref>
<ref id="B31"><mixed-citation publication-type="book"><string-name><surname>R&#225;cz</surname>, <given-names>P.</given-names></string-name> (<year>2025</year>). <source>Word frequency list from the Hungarian Webcorpus</source> (Version 1.0). <publisher-name>Zenodo</publisher-name>. <pub-id pub-id-type="doi">10.5281/zenodo.17508385</pub-id></mixed-citation></ref>
<ref id="B32"><mixed-citation publication-type="journal"><string-name><surname>R&#225;cz</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Beckner</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Hay</surname>, <given-names>J. B.</given-names></string-name>, &#38; <string-name><surname>Pierrehumbert</surname>, <given-names>J. B.</given-names></string-name> (<year>2020</year>). <article-title>Morphological convergence as on-line lexical analogy</article-title>. <source>Language</source>, <volume>96</volume> (<issue>4</issue>), <fpage>735</fpage>&#8211;<lpage>770</lpage>. <pub-id pub-id-type="doi">10.1353/lan.2020.0061</pub-id></mixed-citation></ref>
<ref id="B33"><mixed-citation publication-type="journal"><string-name><surname>R&#225;cz</surname>, <given-names>P.</given-names></string-name>, &#38; <string-name><surname>Luk&#225;cs</surname>, <given-names>&#193;.</given-names></string-name> (<year>2024</year>). <article-title>Lexical and social effects on the learning and integration of inflectional morphology</article-title>. <source>Cognitive Science</source>, <volume>48</volume> (<issue>8</issue>), <elocation-id>e13483</elocation-id>. <pub-id pub-id-type="doi">10.1111/cogs.13483</pub-id></mixed-citation></ref>
<ref id="B34"><mixed-citation publication-type="journal"><string-name><surname>Roberts</surname>, <given-names>G.</given-names></string-name> (<year>2010</year>). <article-title>An experimental study of social selection and frequency of interaction in linguistic diversity</article-title>. <source>Interaction Studies</source>, <volume>11</volume> (<issue>1</issue>), <fpage>138</fpage>&#8211;<lpage>159</lpage>. <pub-id pub-id-type="doi">10.1075/is.11.1.06rob</pub-id></mixed-citation></ref>
<ref id="B35"><mixed-citation publication-type="journal"><string-name><surname>Schumacher</surname>, <given-names>R. A.</given-names></string-name>, &#38; <string-name><surname>Pierrehumbert</surname>, <given-names>J. B.</given-names></string-name> (<year>2021</year>). <article-title>Familiarity, consistency, and systematizing in morphology</article-title>. <source>Cognition</source>, <volume>212</volume>, <elocation-id>104512</elocation-id>. <pub-id pub-id-type="doi">10.1016/j.cognition.2020.104512</pub-id></mixed-citation></ref>
<ref id="B36"><mixed-citation publication-type="book"><string-name><surname>Sipt&#225;r</surname>, <given-names>P.</given-names></string-name>, &#38; <string-name><surname>T&#246;rkenczy</surname>, <given-names>M.</given-names></string-name> (<year>2000</year>). <source>The phonology of Hungarian</source>. <publisher-name>Oxford University Press</publisher-name>. <pub-id pub-id-type="doi">10.1093/oso/9780198238416.001.0001</pub-id></mixed-citation></ref>
<ref id="B37"><mixed-citation publication-type="journal"><string-name><surname>Stivers</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Brown</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Englert</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Hayashi</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Heinemann</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Hoymann</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Rossano</surname>, <given-names>F.</given-names></string-name>, <string-name><surname>De Ruiter</surname>, <given-names>J. P.</given-names></string-name>, <string-name><surname>Yoon</surname>, <given-names>K.-E.</given-names></string-name>, et al. (<year>2009</year>). <article-title>Universals and cultural variation in turn-taking in conversation</article-title>. <source>Proceedings of the National Academy of Sciences</source>, <volume>106</volume> (<issue>26</issue>), <fpage>10587</fpage>&#8211;<lpage>10592</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0903616106</pub-id></mixed-citation></ref>
<ref id="B38"><mixed-citation publication-type="book"><string-name><surname>Szmrecsanyi</surname>, <given-names>B.</given-names></string-name> (<year>2006</year>). <source>Morphosyntactic persistence in spoken English: A corpus study at the intersection of variationist sociolinguistics, psycholinguistics, and discourse analysis</source>. <publisher-name>Walter de Gruyter</publisher-name>. <pub-id pub-id-type="doi">10.1515/9783110197808</pub-id></mixed-citation></ref>
<ref id="B39"><mixed-citation publication-type="journal"><string-name><surname>T&#243;th</surname>, <given-names>E.</given-names></string-name> (<year>2007</year>). <article-title>The imperative and the subjunctive proper in Hungarian</article-title>. <source>Sprachtheorie und germanistische Linguistik</source>, <volume>17</volume> (<issue>2</issue>), <fpage>125</fpage>&#8211;<lpage>145</lpage>.</mixed-citation></ref>
<ref id="B40"><mixed-citation publication-type="journal"><string-name><surname>Vehtari</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Gelman</surname>, <given-names>A.</given-names></string-name>, &#38; <string-name><surname>Gabry</surname>, <given-names>J.</given-names></string-name> (<year>2017</year>). <article-title>Practical Bayesian model evaluation using leave-one-out cross-validation and WAIC</article-title>. <source>Statistics and Computing</source>, <volume>27</volume> (<issue>5</issue>), <fpage>1413</fpage>&#8211;<lpage>1432</lpage>. <pub-id pub-id-type="doi">10.1007/s11222-016-9696-4</pub-id></mixed-citation></ref>
<ref id="B41"><mixed-citation publication-type="journal"><string-name><surname>Weiner</surname>, <given-names>E. J.</given-names></string-name>, &#38; <string-name><surname>Labov</surname>, <given-names>W.</given-names></string-name> (<year>1983</year>). <article-title>Constraints on the agentless passive</article-title>. <source>Journal of Linguistics</source>, <volume>19</volume> (<issue>1</issue>), <fpage>29</fpage>&#8211;<lpage>58</lpage>. <pub-id pub-id-type="doi">10.1017/S0022226700007441</pub-id></mixed-citation></ref>
<ref id="B42"><mixed-citation publication-type="book"><string-name><surname>Wickham</surname>, <given-names>H.</given-names></string-name> (<year>2016</year>). <source>ggplot2: Elegant graphics for data analysis</source>. <publisher-name>Springer-Verlag</publisher-name>. <pub-id pub-id-type="doi">10.1007/978-3-319-24277-4</pub-id></mixed-citation></ref>
</ref-list>
</back>
</article>