<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.0 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.0/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.0" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">2397-1835</journal-id>
<journal-title-group>
<journal-title>Glossa: a journal of general linguistics</journal-title>
</journal-title-group>
<issn pub-type="epub">2397-1835</issn>
<publisher>
<publisher-name>Ubiquity Press</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5334/gjgl.180</article-id>
<article-categories>
<subj-group>
<subject>Research</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Reciprocal expressions and the Maximal Typicality Hypothesis</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<contrib-id contrib-id-type="orcid">http://orcid.org/0000-0003-3660-7333</contrib-id>
<name>
<surname>Poortman</surname>
<given-names>Eva B.</given-names>
</name>
<email>e.b.poortman@uu.nl</email>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Struiksma</surname>
<given-names>Marijn E.</given-names>
</name>
<email>m.struiksma@uu.nl</email>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Kerem</surname>
<given-names>Nir</given-names>
</name>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Friedmann</surname>
<given-names>Naama</given-names>
</name>
<xref ref-type="aff" rid="aff-3">3</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Winter</surname>
<given-names>Yoad</given-names>
</name>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>Utrecht University, Trans 10, 3512 JK, Utrecht, NL</aff>
<aff id="aff-2"><label>2</label>Google Israel Ltd., Yigal Alon 98, Tel Aviv 6789141, IL</aff>
<aff id="aff-3"><label>3</label>Tel Aviv University, Tel Aviv 69978, IL</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2018-02-02">
<day>02</day>
<month>02</month>
<year>2018</year>
</pub-date>
<pub-date pub-type="collection">
<year>2018</year>
</pub-date>
<volume>3</volume>
<issue>1</issue>
<elocation-id>18</elocation-id>
<history>
<date date-type="received" iso-8601-date="2016-06-22">
<day>22</day>
<month>06</month>
<year>2016</year>
</date>
<date date-type="accepted" iso-8601-date="2017-09-19">
<day>19</day>
<month>09</month>
<year>2017</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2018 The Author(s)</copyright-statement>
<copyright-year>2018</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.glossa-journal.org/articles/10.5334/gjgl.180/"/>
<abstract>
<p>In two experiments, we study the effects of verb concepts on the interpretation of reciprocal expressions in Dutch and Hebrew. One experiment studies Hebrew to test a previous account, the Strongest Meaning Hypothesis, which suggests that listeners resolve ambiguity in reciprocal sentences using the logically strongest meaning that is consistent with the context. The results challenge this proposal, as participants often adopt a weaker meaning than what the Strongest Meaning Hypothesis expects. We propose that these results reflect the sensitivity of reciprocal quantifiers to verb concepts, which is modelled by a new principle, the <italic>Maximal Typicality Hypothesis</italic> (MTH). For any given reciprocal sentence, the MTH specifies a <italic>core situation</italic>: the maximal situation that is also maximally typical for the verb concept. The MTH predicts reciprocal sentences to be maximally acceptable in the core situation and, under certain conditions, in situations that contain it, but substantially less acceptable in other situations. To test this prediction, we conducted a two-part experiment among Dutch speakers: (a) a membership test that ranks typicality preferences with different verbs; (b) a truth-value judgement test with reciprocal sentences containing these verbs. The results show that the typical number of patients per agent varies between verbs, with a significant effect of these preferences on reciprocal quantification: the stronger the verb concept&#8217;s bias is for one-patient situations, the weaker is the interpretation of reciprocal sentences containing it. These results support the MTH as a basis for a general theory of reciprocal quantification.</p>
</abstract>
<kwd-group>
<kwd>concepts</kwd>
<kwd>typicality effects</kwd>
<kwd>reciprocity</kwd>
<kwd>Strongest Meaning Hypothesis (SMH)</kwd>
<kwd>Maximal Typicality Hypothesis (MTH)</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec>
<title>1 Introduction</title>
<p>A well-known fact about reciprocal expressions like <italic>each other</italic> and <italic>one another</italic> is their sensitivity to linguistic and non-linguistic contextual parameters. Of these parameters, there is a special place for the meaning of lexical predicates with which reciprocals appear. When using a complex reciprocal verb phrase like <italic>admire each other</italic> or <italic>chase each other</italic>, the first linguistic factor that affects reciprocity is the verb&#8217;s meaning. However, as with other content words, meanings of verbs like <italic>admire</italic> or <italic>chase</italic> are notoriously hard to define. One of the challenges for a semantic theory of reciprocity stems directly from this fuzziness. Despite the considerable efforts that have been invested in studying reciprocal expressions, all previous works analyze their quantificational effects as separate from the fuzzy aspects of predicate meaning. This separation between predicate meanings and quantificational processes has led to considerable empirical shortcomings of even the most precise theories of reciprocals (<xref ref-type="bibr" rid="B2">Dalrymple et al. 1998</xref>; <xref ref-type="bibr" rid="B20">Sabato and Winter 2012</xref>). As we will show, there are intriguing semantic regularities in the way quantification with reciprocal expressions is affected by verb concepts. These regularities challenge previous approaches, and lead us to a new, experimentally informed, theory of reciprocals. The theory that we propose accounts for the facts that Dalrymple et al.&#8217;s proposal fails to explain, while preserving some of the empirical predictions and theoretical insights that it aims to articulate. Thereby, the proposed theory extends the analysis of quantificational reciprocity to a previously unstudied terrain between formal semantics and lexical semantics.</p>
<p>A major challenge for previous studies of reciprocals involves pairs of sentences like (1) and (2) below.</p>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(1)</td>
<td>John,&#160;Bill&#160;and&#160;George&#160;know&#160;each&#160;other.</td>
</tr>
<tr>
<td>(2)</td>
<td>John,&#160;Bill&#160;and&#160;George&#160;are&#160;biting&#160;each&#160;other.</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The logical structures of (1) and (2) are identical, consisting of three elements: a reciprocal expression (<italic>each other</italic>), an antecedent set of cardinality two or more (<italic>John, Bill and George</italic>), and a verb concept (<italic>know/bite</italic>). This structure provides speakers with a concise way of describing complex situations. For example, sentence (1) naturally describes a situation with six acquaintance relations:<xref ref-type="fn" rid="n1">1</xref> John knows Bill, Bill knows John, John knows George, George knows John, Bill knows George and George knows Bill. We henceforth refer to any situation that demonstrates this full mutuality between three agents as an <italic>S6 situation</italic>. While the <italic>S6</italic> situation is salient for sentence (1), this is not the case for sentence (2): when hearing sentence (2), we first think of a situation with only three relations, where every man is biting only one other man. We refer to such situations as <italic>S3 situations</italic>.</p>
<p>Since Langendoen&#8217;s (<xref ref-type="bibr" rid="B6">1978</xref>) early work, many studies of reciprocals in formal semantics have presented compelling evidence that reciprocal sentences such as (1) and (2) describe different situations. Some of these works concentrate on the way to derive meanings for reciprocal sentences with different syntactic structures (e.g. <xref ref-type="bibr" rid="B4">Heim, Lasnik and May 1991</xref>; <xref ref-type="bibr" rid="B1">Beck 2001</xref>). Other works focus on the use of contextual and lexical information for selecting between different meanings for reciprocal sentences as in (1) and (2) (<xref ref-type="bibr" rid="B2">Dalrymple et al. 1998</xref>; <xref ref-type="bibr" rid="B20">Sabato &amp; Winter 2012</xref>; <xref ref-type="bibr" rid="B9">Mari 2013</xref>).<xref ref-type="fn" rid="n2">2</xref> This paper develops the second line by experimentally studying the role of lexical information in the formal semantics of reciprocals.</p>
<p>In their influential work, Dalrymple et al. (<xref ref-type="bibr" rid="B2">1998</xref>) show a way to account systematically for the different interpretations of reciprocal sentences. Dalrymple et al. propose that each occurrence of a reciprocal expression denotes one of six different logical operators, encoding different strategies for categorizing situations. For instance, the operator of <italic>Strong Reciprocity</italic> (SR) categorizes situations where every member of the antecedent set is connected by the given relation to every other member. Thus, in sentences like (1) and (2) where the antecedent set consists of three members, applying SR results in selecting <italic>S6</italic> as the only possible situation where the sentence can be truthfully asserted.<xref ref-type="fn" rid="n3">3</xref> Other operators by Dalrymple et al. impose laxer logical requirements, which also admit other situations in addition to <italic>S6</italic>, especially the kind of situation we called <italic>S3</italic>.</p>
<p>To select between the six logical operators they propose, Dalrymple et al. introduce a principle that they call the <italic>Strongest Meaning Hypothesis</italic> (SMH). The selection relies on assumptions about the linguistic and non-linguistic context in which the reciprocal expression is used. For each occurrence of a reciprocal expression in a given context, the SMH selects the operator that results in the <italic>strongest sentential meaning</italic> that is consistent with that context. Other operators are ruled out. For example, in sentence (1), Dalrymple et al.&#8217;s account assumes that the context puts no restrictions on the number of possible acquaintances that any of the three persons may have.<xref ref-type="fn" rid="n4">4</xref> As a result, the SMH selects the strongest sentence meaning, where <italic>each other</italic> is interpreted as the SR operator: every man knows every other man. This analysis predicts that sentence (1) can only be true in <italic>S6</italic>, which is in line with many speakers&#8217; intuitions. Sentence (2) is also successfully analyzed by the SMH, by making the plausible assumption that each individual can only bite one other individual at a time. Taking this assumption on the meaning of <italic>bite</italic> to be part of the context of (2), the SMH cannot select SR for the reciprocal expression in (2), as this would be inconsistent with the contextual information. Consequently, a logically weaker operator than SR is selected (Dalrymple et al.&#8217;s operator of <italic>Intermediate Reciprocity</italic>). This operator correctly analyzes sentence (2) as true in <italic>S3</italic> situations such as the one where John bites Bill, Bill bites George and George bites John, i.e. where each individual only bites one other individual.</p>
<p>We see that based on plausible contextual assumptions, the SMH correctly accounts for the intuitive distinction between sentences (1) and (2). This account relies on the assumption that the verb <italic>bite</italic> restricts the number of possible patients per agent, whereas the verb <italic>know</italic> does not. More generally, in all the cases that Dalrymple et al. discuss, their analysis assumes that the context categorically restricts the number of patients that agents may possibly have simultaneously, or the number of agents that patients may be connected to. This assumption is needed in order for the SMH to allow reciprocal meanings that are weaker than Strong Reciprocity. In sentence (2), this assumption is innocuous enough: it looks reasonable to assume that ordinary contexts categorically rule out situations where some agent bites more than one patient simultaneously. However, with many other sentences, this kind of categorical judgement is questionable. For instance, let us consider sentence (3) below.</p>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(3)</td>
<td>John,&#160;Bill&#160;and&#160;George&#160;are&#160;pinching&#160;each&#160;other.</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In terms of the situations that it supports, sentence (3) does not differ much from sentence (2): intuitively, both sentences seem equally true in <italic>S3</italic> situations. However, despite this similarity in the sentential interpretations of (2) and (3), the predicates <italic>pinch</italic> and <italic>bite</italic> are quite different in terms of the meaning restrictions that they induce on possible contexts. While in natural contexts, biting two people simultaneously seems impossible, pinching two patients simultaneously is physically possible. As a result, in normal contexts where people use both hands, the SMH would select SR as the meaning of the reciprocal in (3), hence the SMH expects the sentence to only be acceptable in the <italic>S6</italic> situation. This prediction of the SMH seems problematic: intuitively, we believe that sentence (3) should be judged acceptable in <italic>S3</italic>. As we will see, this intuition is supported by our experimental findings.</p>
<p>When addressing this challenge for the SMH, we might choose one of two opposing lines:</p>
<list list-type="order">
<list-item><p>In an attempt to salvage the SMH, we might assume that for many speakers, the context does in fact put restrictions on the number of possible patients that may be pinched simultaneously, similarly to how the number of possible patients is restricted for <italic>bite</italic> in (2). In such contexts, the SMH would again select a weaker logical operator than SR. This would license sentence (3) in <italic>S3</italic> situations.</p></list-item>
<list-item><p>An opposite approach would be to conjecture that reciprocals allow weaker meanings than SR for all sentences. One candidate for such a meaning might be the one that Langendoen (<xref ref-type="bibr" rid="B6">1978</xref>) calls <italic>Weak Reciprocity</italic>. Without a selection principle like the SMH, the WR operator renders all reciprocal sentences like (1)&#8211;(3) acceptable in <italic>S3</italic> situations.</p></list-item>
</list>
<p>Approach 1 predicts that all speakers who accept sentences like (3) in <italic>S3</italic> situations reject them in <italic>S6</italic>. Approach 2 expects all speakers to also accept sentences like (1) in <italic>S3</italic> situations. As we will see, both predictions turn out to be problematic. To avoid these problems, we need to retain a contextually-informed principle for selecting reciprocal meanings, but the selection process itself must be based on more fine-grained, experimentally testable, assumptions about the lexical semantics of predicates.</p>
<p>Developing this idea, we propose a new selection principle that we call the <italic>Maximal Typicality Hypothesis</italic> (MTH).<xref ref-type="fn" rid="n5">5</xref> Avoiding sharp categorical judgements about concepts as in Dalrymple et al.&#8217;s proposal, the MTH predicts truth-values for reciprocal sentences in correlation to a fuzzy measure of conceptual semantics: the <italic>typicality</italic> of different situations for verb concepts like <italic>know, pinch</italic> or <italic>bite</italic>. For any given reciprocal sentence, the MTH designates a so-called <italic>core situation</italic>: a situation that contains enough relations between the agents in the sentence to satisfy reciprocity. What we define as &#8220;enough relations&#8221; depends on typicality information about the verb concept. In sentence (1), only the maximum of six acquaintance relations between the men would be considered enough, since the concept <italic>know</italic> imposes no typicality restrictions with respect to the number of simultaneous patients per agent. Thus, the core situation for (1) is the maximal one, where every man knows every other man (<italic>S6</italic>). This is the only situation in which sentence (1) is expected to be fully acceptable. By contrast, in (2) and (3), both verb concepts have clear typicality preferences with respect to the number of simultaneous patients per agent. Having six relations (i.e. two per agent) as in <italic>S6</italic> would be against the typicality preferences of the verb concept (in the case of <italic>pinch</italic>), or would not even be considered an instance of the verb concept (in the case of <italic>bite</italic>). In such cases, the core situation that the MTH selects is the maximal situation with no more than one patient per agent, hence <italic>S3</italic> is selected as core situation for (2) and (3). Accordingly, the MTH expects both (2) and (3) to be true and fully acceptable in <italic>S3</italic>.</p>
<p>This qualitative analysis is tested in two experiments, which are reported in Sections 3 and 4. Furthermore, the MTH makes novel and more precise quantitative predictions about the relations between reciprocity and typicality. Since acceptability of reciprocal sentences is predicted on the basis of typicality preferences with verb concepts, the MTH expects that the lower the typicality of situations with two patients per agent, the higher the prominence of readings that are weaker than SR. The experiment reported in Section 4 tests this predicted correlation.</p>
<p>The paper is structured as follows. Section 2 introduces the MTH in detail, as an alternative to the SMH. Sections 3 and 4 report two experiments that tested the afore-mentioned predictions of the MTH against those of the SMH. Section 5 concludes the paper.</p>
</sec>
<sec>
<title>2 The Maximal Typicality Hypothesis</title>
<p>This section introduces the Maximal Typicality Hypothesis as a solution to the problems that are raised by the different acceptability patterns for sentences like (1)&#8211;(3). After some background on typicality in subsection 2.1, the MTH is introduced in section 2.2. Further, subsections 2.3&#8211;2.5 explain our method for testing the MTH.</p>
<sec>
<title>2.1 Concept typicality</title>
<p>In terms of theories of mental concepts (<xref ref-type="bibr" rid="B8">Margolis &amp; Laurence 1999</xref>), Dalrymple et al.&#8217;s formulation of the SMH only takes into account &#8220;sharp&#8221; aspects of the meaning of verb concepts such as <italic>know, bite</italic>, and <italic>pinch</italic>: whether certain situations &#8211; e.g. those compatible with SR &#8211; are possible or impossible in a given context. As mentioned above, such sharp distinctions do not easily account for the intuitive contrasts between sentences like (1), (2), and (3). To overcome this shortcoming of the SMH, the proposed Maximal Typicality Hypothesis revises the SMH by on the basis of the typicality preferences of verb concepts.</p>
<p>Since the 1970&#8217;s, a host of psychological studies has shown that subjects consistently rank some instances of a one-place predicate concept as more typical than others (e.g. <xref ref-type="bibr" rid="B18">Rosch 1973</xref>; <xref ref-type="bibr" rid="B22">Smith et al. 1974</xref>; <xref ref-type="bibr" rid="B19">Rosch and Mervis 1975</xref>). For example, besides being able to categorize sparrows and ostriches within the <italic>bird</italic> category, and koalas and crocodiles outside of it, subjects also distinguish between members of a category: e.g. when subjects are asked to rank bird instances, sparrows are judged as more typical for the concept <italic>bird</italic> than ostriches. These rankings correlate with other measures of typicality, such as categorization speed (more typical instances are categorized faster than less typical ones) and error rate (more typical instances lead to fewer categorization errors than less typical ones). Throughout this paper, we will use the term &#8216;typicality effect&#8217; to refer to this basic behavioral phenomenon about categorization.</p>
<p>While many nouns categorize simple entities, verbs categorize more complex <italic>situations</italic>: events and states containing different entities as participants. As we will show, verb concepts exhibit typicality effects with situations, similarly to the typicality effects that noun concepts show with entities. For reciprocal sentences and verb concepts, taking typicality into account means changing perspectives about Dalrymple et al.&#8217;s notion of context. In addition to the definitional aspects that the SMH considers (can a given situation be categorized as an instance of a verb concept <italic>X</italic>?), our proposed account also has recourse to aspects of typicality (what preferences between situations does a verb concept <italic>X</italic> induce?). This allows us to take into account more factors that affect the interpretation of reciprocal sentences. As we saw, the sharp distinction that the SMH makes between possible and impossible situations forces it to choose SR as the meaning of the reciprocal expression in both sentences (1) and (3). This is intuitively questionable for (3). Under the theory we propose, the interpretation of the two sentences differs due to differences in typicality information between the concepts <italic>know</italic> and <italic>pinch</italic>. More generally, our theory uses experimental evidence on typicality to make predictions about the interpretations of reciprocal sentences. As we show, this leads to more fine-grained predictions, which make the correct distinctions between sentences like (1), (2) and (3).</p>
</sec>
<sec>
<title>2.2 The MTH: Connecting typicality with the interpretation of reciprocals</title>
<p>In the formal semantic literature on reciprocals, there are two general approaches that can be discerned. One approach, as in Langendoen (<xref ref-type="bibr" rid="B6">1978</xref>), Dalrymple et al. (<xref ref-type="bibr" rid="B2">1998</xref>), or Beck (<xref ref-type="bibr" rid="B1">2001</xref>), assumes systematic ambiguity of reciprocal expressions. Another approach, suggested by Roberts (<xref ref-type="bibr" rid="B17">1987</xref>), is that reciprocals are truth-conditionally vague, similarly to quantificational words like <italic>many, most</italic> or <italic>enough</italic>. The important contribution of Dalrymple et al.&#8217;s SMH is its ability to analyze ambiguity resolution systematically based on contextual information. Sabato and Winter (<xref ref-type="bibr" rid="B20">2012</xref>) propose to use Dalrymple et al.&#8217;s insights without assuming that reciprocals are ambiguous, but by letting them have a general functional meaning that takes lexical or contextual knowledge about predicate meanings as its argument. This semantic strategy is closer to Roberts&#8217; proposal in its aim to avoid ambiguity of reciprocals. However, like Dalrymple et al., Sabato and Winter also assume categorical distinctions between possible and impossible predicate denotations. What these previous approaches lack is a fine-grained notion of predicate meanings which takes typicality preferences into account, and uses it to analyze speaker judgements on sentences like (3). The present proposal follows Roberts&#8217;s initial approach by incorporating typicality preferences into a theory that develops Sabato and Winter&#8217;s version of the SMH.</p>
<p>The proposed Maximal Typicality Hypothesis (MTH) uses information about typicality of different situations with respect to a given verb concept <italic>P</italic>. On the basis of this information on <italic>P</italic>, the MTH singles out one situation as the core situation for reciprocal sentences containing <italic>P</italic>. This core situation is the basis for speakers&#8217; interpretation of reciprocal sentences.<xref ref-type="fn" rid="n6">6</xref> Above we intuitively described the core situation as the situation that contains <italic>enough</italic> relations between the agents to fully satisfy reciprocity. In terms of reciprocity, for sentence (1) with the verb <italic>know</italic>, situation <italic>S6</italic> has enough relations, whereas for sentence (3) with the verb <italic>pinch</italic>, already <italic>S3</italic> has enough relations. To be more precise, we need to define what we mean by <italic>enough for reciprocity</italic>, and how the two sentences differ in this respect. According to the MTH, the difference between (1) and (3) follows from the observed typicality difference between the verbs. With the verb <italic>pinch</italic> we cannot add relations to <italic>S3</italic> without reaching a situation where one agent pinches two patients. Such a situation would be atypical for the verb <italic>pinch</italic>. Because we cannot add relations to <italic>S3</italic> without a reduction in its typicality for <italic>pinch</italic>, we consider <italic>S3</italic> to have enough relations for sentence (3). By contrast, with the verb <italic>know</italic>, we can add relations to <italic>S3</italic> without any change in typicality for the verb concept. Thus, <italic>S3</italic> does not attain enough relations for sentence (1). In technical terms, we describe this difference by observing that <italic>S3</italic> is the <italic>maximal situation</italic> among the most typical situations for the verb <italic>pinch</italic>. By contrast, <italic>S3</italic> is not maximal among the most typical situations for the verb <italic>know</italic>. Accordingly, the MTH selects <italic>S3</italic> as the core situation for sentence (3), but not for sentence (1). In general, this process of selecting the core situation is defined below.</p>
<disp-quote>
<p><bold>Maximal Typicality Hypothesis (MTH):</bold> For a reciprocal sentence with a verb concept <italic>P</italic> in the scope of the reciprocal expression, situation <italic>S<sub>C</sub></italic> is the <italic>core situation</italic> for the sentence iff <italic>S<sub>C</sub></italic> is maximal among the situations that are most typical for <italic>P</italic>.<xref ref-type="fn" rid="n7">7</xref></p>
</disp-quote>
<p>The MTH assumes that similar to noun concepts, verb concepts invoke typicality judgements on situations.<xref ref-type="fn" rid="n8">8</xref> Formally (<xref ref-type="bibr" rid="B23">Winter 2017</xref>): the MTH defines the core situation as any maximal situation <italic>S<sub>C</sub></italic> among the situations <italic>S</italic> that attain a local maximum of the function <italic>TYP<sub>P</sub>(S)</italic> &#8211; the typicality of <italic>S</italic> for the predicate <italic>P</italic>. This typicality is a variable whose values can be experimentally estimated. Once we know which situations are most typical for a given verb concept, the MTH predicts the core situation for a reciprocal sentence containing that verb. The key for selecting the core situation among the most typical situations is the notion of <italic>maximal situation</italic>. This notion relies on our ability to order situations according to containment relations between them. For example, a situation like <italic>S6</italic>, where every agent acts on each of the two other agents, properly contains any <italic>S3</italic> situation. To illustrate this notion, it is helpful to consider the diagrams in Figure <xref ref-type="fig" rid="F1">1</xref>.</p>
<fig id="F1">
<label>Figure 1</label>
<caption><p>Three situations with three individuals.</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66643/"/>
</fig>
<p>The arrows in Figure <xref ref-type="fig" rid="F1">1</xref> represent directed actions like <italic>pinch</italic> between agents. Thus, Figures <xref ref-type="fig" rid="F1">1B</xref> and <xref ref-type="fig" rid="F1">1C</xref> represent <italic>S3</italic> and <italic>S6</italic> respectively. Figure <xref ref-type="fig" rid="F1">1A</xref> represents another kind of situation, to which we refer as <italic>S2</italic>. Since situation <italic>S6</italic> contains <italic>S3</italic>, and <italic>S3</italic> contains <italic>S2</italic>, we conclude that <italic>S6</italic> is maximal among the three situations in Figure <xref ref-type="fig" rid="F1">1</xref>. In sentence (1), we assume that <italic>S6, S3</italic> and <italic>S2</italic> are of equal typicality for the verb <italic>know</italic> (at least as far as the number of patients is concerned, see section 2.3 below). Among these three situations <italic>S6</italic> is the maximal, hence it is the one that is selected by the MTH. By contrast, in sentences (2) and (3), we assume that <italic>S2</italic> and <italic>S3</italic> are more typical for the verbs <italic>bite</italic> and <italic>pinch</italic> than <italic>S6</italic> (furthermore, <italic>S6</italic> may even be impossible with <italic>bite</italic>). The maximal situation among the most typical situations is now <italic>S3</italic>, and accordingly, this is the situation that is selected by the MTH as the core situation.</p>
<p>This selection of the core situation for reciprocal sentences like (1), (2) or (3) is the basis for explaining the acceptability pattern we observe with these sentences in all situations, including those that are not the core situation. Reciprocal sentences are always expected to be highly acceptable in the core situation. In addition to this core situation, there are two kinds of situations that we should consider:<xref ref-type="fn" rid="n9">9</xref></p>
<list list-type="alpha-lower">
<list-item><p><italic>Situations that are properly contained in the core situation</italic>: In those situations, reciprocal sentences are expected to be less acceptable than in the core situation, with decreasing acceptability the fewer relations there are. For example, the MTH defines <italic>S6</italic> as the core situation for sentence (1). Accordingly, sentence (1) is predicted to be less acceptable in <italic>S3</italic> than in <italic>S6</italic>, and less acceptable in <italic>S2</italic> than in <italic>S3</italic>.<xref ref-type="fn" rid="n10">10</xref></p></list-item>
<list-item><p><italic>Situations that properly contain the core situation</italic>: In those situations, there are &#8220;more than enough&#8221; relations to support reciprocity. Whether such situations are relevant for actual use of a given reciprocal sentence is determined by the felicity that speakers assign to them. For example, both sentences (2) and (3) are predicted to be highly acceptable in their core situation, <italic>S3</italic>. As for <italic>S6</italic> situations, judgements are affected by how felicitous situation <italic>S6</italic> is judged to be as a possible instance of the verb. Situation <italic>S6</italic> is usually judged to be a possible instance of the verb <italic>pinch</italic>. Accordingly, speakers are expected to judge sentence (3) as highly acceptable in <italic>S6</italic>. By contrast, <italic>S6</italic> situations are unlikely to be judged as acceptable for the concept <italic>bite</italic> to begin with. Accordingly, such situations are harder to use with sentence (2), and are often judged as unacceptable for this sentence.<xref ref-type="fn" rid="n11">11</xref></p></list-item>
</list>
</sec>
<sec>
<title>2.3 Approximating typicality of complex situations</title>
<p>In order to test the MTH and its use with different situations, we first need information regarding the typicality of situations like <italic>S2, S3</italic> and <italic>S6</italic> for different verb concepts. The reason we consider situations like <italic>S6</italic> as atypical for the verb <italic>pinch</italic> is because agents in it pinch more than one patient simultaneously. Using this kind of typicality information, the predictions of the MTH are tested against acceptability judgements on reciprocal sentences in situations like <italic>S2, S3</italic> and <italic>S6</italic>. But how can we systematically gather experimental information about the typicality of situations like <italic>S2, S3</italic> and <italic>S6</italic> for different verb concepts? As was discussed in subsection 2.1, previous studies have measured typicality effects using standard tasks such as categorizing instances or ranking them with respect to how typical they are. Those works have mostly measured typicality effects for concepts expressed by nouns. For situations like <italic>S2, S3</italic> and <italic>S6</italic>, we approximate typicality by looking at their sub-situations in relation to typicality preferences of the verb. Consider for example the two situations in Figure <xref ref-type="fig" rid="F2">2</xref>, which depict <italic>S3</italic> and <italic>S6</italic> situations with the verb <italic>pinch</italic>.</p>
<fig id="F2">
<label>Figure 2</label>
<caption><p>Two possible situations for the verb <italic>pinch</italic>: <italic>S3</italic> (A) and <italic>S6</italic> (B).</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66644/"/>
</fig>
<p>Both situations show three women in pinching activities. In situation A, every woman is pinching only one other woman (<italic>S3</italic>), while in situation B, every woman is pinching every other woman (<italic>S6</italic>). Because of the agents using both hands in situation B, this situation is expected to be less typical for the verb <italic>pinch</italic> than situation A. When testing this expectation, we refer to the number of patients acted upon by a given agent as <italic>patient cardinality</italic>. For instance, in Figure <xref ref-type="fig" rid="F2">2</xref>, we describe the difference between the two situations by saying that the patient cardinality value is <italic>1</italic> (=one patient per agent) in Figure <xref ref-type="fig" rid="F2">2A</xref>, but <italic>2</italic> (=two patients per agent) in Figure <xref ref-type="fig" rid="F2">2B</xref>. This reduces the comparison between the two situations in Figure <xref ref-type="fig" rid="F2">2</xref> to a comparison between the two situations in Figure <xref ref-type="fig" rid="F3">3</xref>. Our typicality experiments test which of the situations in Figure <xref ref-type="fig" rid="F3">3</xref> is preferred as a more typical instance of the verb <italic>pinch</italic>. Based on the preference showed for Figure <xref ref-type="fig" rid="F3">3A</xref> over Figure <xref ref-type="fig" rid="F3">3B</xref>, we infer that <italic>S3</italic> situations (Figure <xref ref-type="fig" rid="F1">1B</xref>) are more typical for the verb <italic>pinch</italic> than <italic>S6</italic> situations (Figure <xref ref-type="fig" rid="F1">1C</xref>). This information is used for evaluating the MTH.</p>
<fig id="F3">
<label>Figure 3</label>
<caption><p>Two instances of the verb <italic>pinch</italic>: a one-patient situation (A) and a two-patient situation (B).</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66645/"/>
</fig>
</sec>
<sec>
<title>2.4 Predictions of the MTH</title>
<p>Based on the measurement of typicality effects described above in subsection 2.3, let us illustrate the precise empirical predictions of the MTH for sentences (1)&#8211;(3). For the verb <italic>know</italic> in (1), our experiments show that there is no typicality preference between simple situations with different patient cardinality. This means that there is no difference in the representativeness of different instances of <italic>knowing</italic> for the concept, at least in terms of how many patients each agent knows. For instance, a state in which a person knows one other person is just as typical an instance of <italic>know</italic> as a state in which a person knows two other people. Using this measurement, we extrapolate that there is also no difference in typicality between reciprocal situations that differ merely in terms of patient cardinality (at least in terms of one patient vs. two). This means that the reciprocal situations <italic>S2, S3</italic> and <italic>S6</italic> are equally typical for the verb concept <italic>know</italic>. With this extrapolation about typicality of candidate situations, the MTH predicts that the core situation that (1) describes is <italic>S6</italic>: the maximal situation among the three situations <italic>S2, S3</italic> and <italic>S6</italic>. We thus predict that (1) is fully acceptable in <italic>S6</italic>. This agrees with the prediction of the SMH that the reciprocal in (1) is interpreted as the SR operator. Furthermore, we predict (1) to be less acceptable in <italic>S3</italic>, and even less so in <italic>S2</italic>, because these situations are properly contained in the core situation <italic>S6</italic>. This is a more fine-grained prediction than the truth/falsity prediction of the SMH, which expects (1) to be equally judged as false in both <italic>S2</italic> and <italic>S3</italic>.</p>
<p>In contrast to the verb <italic>know</italic>, the verb <italic>bite</italic> in (2) does show a typicality effect with respect to patient cardinality. In particular, instances of the concept <italic>bite</italic> that have one patient per agent are judged as being more typical than instances that have two patients per agent simultaneously.<xref ref-type="fn" rid="n12">12</xref> We extrapolate from this that in the most typical reciprocal situations for the verb <italic>bite</italic>, there is no more than one patient per agent. The MTH predicts that the core situation described by sentence (2) is the maximal one among those situations. This is the situation in which each agent bites exactly one patient: <italic>S3</italic>. Adding more <italic>biting</italic> relations to such a situation would result in a situation that is not among the most typical situations. Thus, we predict (2) to be fully acceptable in <italic>S3</italic>, in line with the predictions of the SMH. Furthermore, we predict (2) to be unacceptable in <italic>S6</italic>: although the sentence is formally expected to be true in such a hypothetical situation, this scenario is unlikely to be a possible instance of the concept <italic>bite</italic> (see note 12). Finally, sentence (2) is expected to be less acceptable in <italic>S2</italic> than in the core situation <italic>S3</italic>, because <italic>S2</italic> is properly contained in this core situation.</p>
<p>The example that most clearly distinguishes the predictions of the MTH and the SMH is sentence (3). In this case, the core situation that is specified by the MTH does not coincide with the strongest possible interpretation that is selected by the SMH. As we saw, the SR meaning, where every boy is pinching every other boy, is physically possible for sentence (3). Therefore, according to the SMH, SR is the only possible reading for the reciprocal expression, and <italic>S6</italic> is expected to be the only situation in which sentence (3) is true. By contrast, the MTH predicts sentence (3) to be widely accepted in situation <italic>S3</italic>. As we will experimentally show, the concept <italic>pinch</italic> shows the expected typicality effect with respect to patient cardinality: a situation in which one agent pinches one patient is a more typical instance of <italic>pinch</italic> than a situation in which one agent pinches two patients simultaneously. In this sense, the verb <italic>pinch</italic> is similar to the verb <italic>bite</italic>, and we again extrapolate that the most typical situations for the verb <italic>pinch</italic> are those in which there is at most one patient per agent. From these situations, the MTH selects the maximal one, <italic>S3</italic>, as the core situation that the sentence describes. Accordingly, we expect sentence (3) to be fully acceptable in <italic>S3</italic>. In addition, since situation <italic>S6</italic> properly contains the core situation <italic>S3</italic>, we expect the sentence to be true in <italic>S6</italic>. Unlike <italic>bite</italic> in (2), the concept <italic>pinch</italic> allows <italic>S6</italic> as a possible, though atypical, instance of the concept, and hence we predict (3) to be acceptable in <italic>S6</italic>. The expectation that reciprocal sentences like (3) are judged as acceptable in both <italic>S3</italic> and <italic>S6</italic> clearly distinguishes the predictions of the MTH from those of the SMH.</p>
<p>In order to test these predictions, we tested whether and how typicality preferences in terms of patient cardinality predict reciprocal interpretations using the MTH. As we will show, judgements on sentences like (3) reveal a systematic advantage of the MTH over the SMH.</p>
</sec>
<sec>
<title>2.5 Overview of experimental investigation</title>
<p>The experimental work presented in this paper had two goals. Firstly, we aimed to test interpretations of sentences like <italic>the girls are pinching each other</italic>, where the predictions of the SMH appear to be too strong. Secondly, we aimed to test whether the MTH makes the right predictions, with a focus on correlations between typicality preferences and reciprocal interpretation. The two experiments we report were designed to achieve these aims.</p>
<p>Experiment 1 tested the predictions of the SMH among Hebrew speakers. The experiment focused on action verbs in Hebrew like <italic>cavat</italic> (&#8216;pinch&#8217;) that are expected to show a typicality preference for one-patient situations over two-patient situations. The materials of Experiment 1 were constructed based on a pretest that measured patient cardinality preferences for a large group of action verbs. The aim of the pretest was to select verbs for the main experiment that most clearly prefer one-patient situations over two-patient situations. The main experiment itself tested which situations are preferred for reciprocal sentences containing those action verbs that showed a preference for one-patient situations. Such a forced-choice task was used because challenging the SMH requires making sure that more than one situation is considered possible in the context of the sentence. For each sentence, participants chose between two realistically depicted situations with three human agents each: situation <italic>S6</italic>, in which every individual acts on every other individual (e.g. Figure <xref ref-type="fig" rid="F2">2B</xref>), and situation <italic>S3</italic> &#8211; one where every individual acts on exactly one other individual and is acted on by exactly one other individual (e.g. Figure <xref ref-type="fig" rid="F2">2A</xref>). The SMH predicts all sentences to be true in <italic>S6</italic> and false in <italic>S3</italic>, hence for <italic>S6</italic> to always be preferred over <italic>S3</italic>.</p>
<p>Next, Experiment 2 systematically tested the predictions of the Maximal Typicality Hypothesis: whether the core situation for a reciprocal sentence is maximal among those situations that are most typical for the verb concept in the reciprocal&#8217;s scope. In order to have a broad variety of test cases for the MTH, Experiment 2 (unlike Experiment 1) included verbs that show different patient cardinality preferences. For logistic reasons, Experiment 2 was conducted with Dutch speakers. However, a pretest of Experiment 2 showed a significant correlation between Dutch and Hebrew in the typicality preferences of the verbs that were tested in Experiment 1. Thus, both experiments tested similar typicality effects in Dutch and Hebrew. In order to test the relation between reciprocal interpretation and typicality, Experiment 2 made use of two parts. Part 1 tested typicality effects for different types of verb concepts in isolation, specifically the preference for one-patient or two-patient situations. Part 2 of the experiment tested reciprocal interpretations, specifically the acceptability of reciprocal sentences in situations <italic>S6</italic> and <italic>S3</italic>. The MTH is evaluated on the basis of a correlation analysis between the two parts. Based on our measurement of typicality in part 1, we expected <italic>S3</italic> situations to be specified as the core situation in part 2 for sentences like <italic>John, Bill and George are pinching each other</italic>. A correlation analysis tested this expected relationship between the preference for a one-patient situation in part 1, and the acceptability of reciprocal sentences in the <italic>S3</italic> situation in part 2. In this way, Experiment 2 tested our MTH-based account and compared it to the predictions of the SMH.</p>
</sec>
</sec>
<sec>
<title>3 Experiment 1: Testing the SMH</title>
<p>Experiment 1 studied Hebrew reciprocal sentences, testing the predictions of the SMH with verbs like <italic>pinch</italic> that are expected to show a preference for one-patient situations. A pretest collected examples of such verbs.</p>
<sec>
<title>3.1 Pretest: Verbs with a one-patient preference</title>
<p>For a set of action verbs, the pretest measured whether there was a preference between situations with one patient per agent vs. situations with two patients per agent.</p>
<sec>
<title>3.1.1 Participants</title>
<p>Fifty-three students from Tel Aviv University and Technion (Israel Institute of Technology) (14 female, age <italic>M</italic> = 24) participated as part of a class. All participants were native Hebrew speakers.</p>
</sec>
<sec>
<title>3.1.2 Materials and procedure</title>
<p>Thirty-two different Hebrew verbs were tested regarding their patient cardinality preferences in a pen-and-paper questionnaire. For twenty-five verbs, we expected a preference for a one-patient situation over a two-patient situation. For seven verbs, we expected a preference for the two-patient situation (see Table A1 in Appendix A). Each test item contained two instances of a verb, illustrated by two drawings showing a one-patient situation and a two-patient situation. Apart from the number of patients, the two situations were as similar as possible. An example is given in Figure <xref ref-type="fig" rid="F3">3</xref>, depicting two instances of the verb <italic>covet</italic>, &#8216;pinch&#8217;. All test items contained a verbal description of the situation in Hebrew: an agent-verb sentence (without a theme) referring to an activity associated with the verb concept, for example <italic>ha-yeled covet</italic>, &#8216;The boy is pinching&#8217;. Participants were instructed to indicate which of the two depicted situations &#8220;better describes the sentence&#8221;.<xref ref-type="fn" rid="n13">13</xref> In all items, due to different gender or age of the agent and patient(s), the subject of the sentence (a boy, girl, man, or woman) visibly referred to the agent in both drawings.</p>
<p>In addition, we included six filler items, which also contained two drawings. Their aim was to avoid automatic answers. The drawings in the filler items differed from one another with respect to a parameter other than the number of patients. For instance, one filler item showed a boy taking a photo in one drawing and merely holding a camera in the other drawing, accompanied by the Hebrew correlate of the sentence &#8220;The boy is taking a photo&#8221;. The task in the filler items was identical to the test items.</p>
<p>There were two versions of the questionnaire, with reversed order of items. In addition, in case the same verb was used for a test item and a filler item, the test item always preceded the filler item in both versions.</p>
</sec>
<sec>
<title>3.1.3 Results</title>
<p>The results of the pretest are given in Table A1 of Appendix A. Different verbs showed different preferences regarding patient cardinality. The proportion of participants who preferred the one-patient drawing over the two-patient drawing ranged from 8% to 90% between verbs. In general, most of the verbs showed a preference for one patient over two, indicating that participants considered the situation in which only one patient was involved as a more typical instance of that verb than the situation with two patients. This is as expected, since we included mostly verbs that we believed would show a one-patient preference.</p>
<p>We used the results of the pretest to select verbs for the main part of the experiment, testing the SMH. The selection procedure is explained in the Materials section of section 3.2.</p>
</sec>
</sec>
<sec>
<title>3.2 Experiment 1 (main part): Interpretation of reciprocal sentences</title>
<p>The main part of Experiment 1 measured the preferred interpretation of Hebrew reciprocal sentences containing verbs that showed a clear preference for one-patient situations in the pretest. We used a forced-choice task to measure which situation is preferred for the given reciprocal sentences: <italic>S6</italic> (where each individual is acting on every other individual) or <italic>S3</italic> (where each individual is acting on exactly one other individual and is acted on by exactly one other individual). In such contexts <italic>S6</italic> is visibly possible, hence according to the SMH, the sentences are expected to be uniformly judged as true in situation <italic>S6</italic> and false in <italic>S3</italic>. Thus, the SMH expects <italic>S6</italic> to be uniformly preferred over <italic>S3</italic>. The experiment tested this null hypothesis.</p>
<sec>
<title>3.2.1 Participants</title>
<p>Fifty students from Tel Aviv University and Technion (Israel Institute of Technology) (20 female, age <italic>M</italic> = 25) participated, either as part of a class or for monetary compensation. All participants were native Hebrew speakers.</p>
</sec>
<sec>
<title>3.2.2 Materials and procedure</title>
<p>Eleven test item sentences were created based on the results of the pretest. We calculated the 95% confidence interval based on the preference for a one-patient situation from the pretest (C.I. 0.526&#8211;0.668). Using the upper boundary of the confidence interval analysis, we selected eleven verbs that showed highest preference for a one-patient situation and easily allowed graphical representation of reciprocal situations. We also included two control verbs that showed high preference for two-patient situations and allowed the required graphical representation. We then included the total of 13 verbs in reciprocal sentences of the form <italic>A, B and C are P each other</italic> (where <italic>A, B</italic> and <italic>C</italic> are proper names and <italic>P</italic> is a verb). An example for such a Hebrew sentence is in (4).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(4)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Dani,</p></list-item>
<list-item><p>Danny,</p></list-item>
</list>
<list list-type="word">
<list-item><p>yoav</p></list-item>
<list-item><p>Yoav</p></list-item>
</list>
<list list-type="word">
<list-item><p>ve&#8209;ro&#8217;i</p></list-item>
<list-item><p>and&#8209;Roy</p></list-item>
</list>
<list list-type="word">
<list-item><p>covtim</p></list-item>
<list-item><p>pinch&#8209;<sc>PRES.SG.MASC</sc>.</p></list-item>
</list>
<list list-type="word">
<list-item><p>ze&#8209;et&#8209;ze.</p></list-item>
<list-item><p>this.<sc>ACC</sc>&#8209;this</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;Danny,&#160;Yoav&#160;and&#160;Roy&#160;pinch/are&#160;pinching&#160;each&#160;other.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The resulting 13 sentences were tested for their interpretation in a forced-choice task with two drawings depicting different situations. For each of the sentences, there were two forced-choice trials. The main trial contained a drawing of <italic>S6</italic> (in which every individual acts on two other individuals) vs. a drawing of <italic>S3</italic> (in which every individual acts only on one other individual). The second trial served as a control, and contained again a drawing of <italic>S3</italic> vs. a drawing of situation <italic>S2</italic> (in which only two of the three individuals act on another individual). Apart from the number of actions, the two drawings were always as similar as possible, and the subject of the sentence clearly referred to the three agents in the drawings. The example trials for sentence (4) are given in Figure <xref ref-type="fig" rid="F4">4</xref>. Participants were instructed to indicate which of the two depicted situations better describes the sentence (see footnote 13).</p>
<fig id="F4">
<label>Figure 4</label>
<caption><p>Stimuli from Experiment 1, examples of forced-choice task: main trial <italic>S6</italic> vs <italic>S3</italic> (A) and control trial <italic>S3</italic> vs <italic>S2</italic> (B) (translated from Hebrew).</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66646/"/>
</fig>
<p>In addition, we included twenty-three filler items which also contained two drawings. These drawings were accompanied by non-reciprocal sentences (e.g. the Hebrew correlate of &#8216;Owen, Luke and Andrew are painting themselves&#8217;). The task was presented as a pen-and-paper questionnaire.</p>
</sec>
<sec>
<title>3.2.3 Results</title>
<p>The results of Experiment 1 demonstrated differences between items. For the main test item (<italic>S6</italic> vs. <italic>S3</italic>), the preference for <italic>S3</italic> ranged from 10% to 67%. Importantly, for most verbs, a substantial proportion of participants preferred <italic>S3</italic> over <italic>S6</italic>. These results are summarized in Table <xref ref-type="table" rid="T1">1</xref>.</p>
<table-wrap id="T1">
<label>Table 1</label>
<caption><p>Experiment 1 results main test item (<italic>S6</italic> vs <italic>S3</italic>).</p></caption>
<table>
<tr>
<th align="left" style="background-color:#f3f4f4;">Verb (translated from Hebrew)</th>
<th align="center" style="background-color:#f3f4f4;">% preference for <italic>S3</italic></th>
</tr>
<tr>
<td align="left" colspan="2"><hr/></td>
</tr>
<tr>
<td align="left">Hug (control verb)</td>
<td align="right">10.2</td>
</tr>
<tr>
<td align="left">Hit</td>
<td align="right">18.8</td>
</tr>
<tr>
<td align="left">Paint</td>
<td align="right">24.5</td>
</tr>
<tr>
<td align="left">Pinch</td>
<td align="right">26.5</td>
</tr>
<tr>
<td align="left">Point at</td>
<td align="right">29.2</td>
</tr>
<tr>
<td align="left">Shake</td>
<td align="right">34.7</td>
</tr>
<tr>
<td align="left">Stab</td>
<td align="right">38.8</td>
</tr>
<tr>
<td align="left">Apply make-up</td>
<td align="right">40.8</td>
</tr>
<tr>
<td align="left">Clean</td>
<td align="right">40.8</td>
</tr>
<tr>
<td align="left">Towel</td>
<td align="right">40.8</td>
</tr>
<tr>
<td align="left">Give a speech (control verb)</td>
<td align="right">42.9</td>
</tr>
<tr>
<td align="left">Scrape</td>
<td align="right">42.9</td>
</tr>
<tr>
<td align="left">Comb</td>
<td align="right">66.7</td>
</tr>
</table>
</table-wrap>
<p>For the control item (<italic>S3</italic> vs <italic>S2</italic>), the vast majority of participants preferred <italic>S3</italic> over <italic>S2</italic> (see Table A2 in Appendix A).</p>
</sec>
</sec>
<sec>
<title>3.3 Discussion</title>
<p>From the pretest, we learned that there are verbs that show a preference for a one-patient situation over a two-patient situation, as expected. We then examined reciprocal sentences that contain those verbs that showed clear one-patient preferences in the pretest. The main part of Experiment 1 tested our expectation that the interpretation of such sentences would be weaker than what the SMH predicts.</p>
<p>The results of Experiment 1 showed a large variability in preferences between <italic>S3</italic> and <italic>S6</italic>. Crucially, for eight out of the thirteen reciprocal sentences that were tested, more than one third of the participants preferred <italic>S3</italic> over <italic>S6</italic>. This result is unexpected by the SMH. The SMH predicts that the meaning of any reciprocal sentence is the strongest one that is consistent with its context. The <italic>S6</italic> situation was explicitly presented to the participants in the forced choice task, hence this situation was possible in the context of the reciprocal sentence.<xref ref-type="fn" rid="n14">14</xref> Since this situation is consistent with SR, and SR is the strongest reciprocal meaning, the SMH predicts SR to be the <italic>only</italic> reading of the reciprocal expression. The <italic>S3</italic> situation is inconsistent with this SR meaning, hence the SMH expects 100 percent preference for <italic>S6</italic> over <italic>S3</italic>. This expectation was not borne out: many subjects preferred <italic>S3</italic> to <italic>S6</italic> as the best situation for the reciprocal sentence. It is difficult for the SMH to explain these preferences. Our hypothesis is that these responses are due to the fact that the tested verbs have a typicality preference for one-patient situations. We hypothesize that from the preference for one-patient situations over two-patient situation one can extrapolate a preference for reciprocal situations with only one-patient relations &#8211; hence the common preference for <italic>S3</italic> over <italic>S6</italic>. Note that the uniform preference of <italic>S3</italic> over <italic>S2</italic> in the control trials suggests that when typicality does not play a role (there is no difference in patient cardinality between these situations) the preferred situation for the sentence is the core situation (<italic>S3</italic>) that the MTH selects. This conclusion is strengthened by the results of Experiment 2.</p>
</sec>
</sec>
<sec>
<title>4 Experiment 2: Testing the MTH</title>
<p>Experiment 2 tested the Maximal Typicality Hypothesis as a remedy to the problems we encountered for the SMH. The experiment contained two parts. Part 1 tested typicality differences between verb concepts, specifically the preference between one-patient situations vs. two-patient situations. Part 2 tested reciprocal interpretations, specifically whether reciprocal sentences are accepted in <italic>S6</italic> and <italic>S3</italic> situations. When a speaker prefers a one-patient situation for a given verb, the MTH expects the core situation for the reciprocal sentence to be <italic>S3</italic>, with <italic>S6</italic> as another possible situation. Conversely, with verbs that do not show a preference for one-patient situations, the core reciprocal situation is expected to be <italic>S6</italic>. Accordingly, for a representative sample of verbs, the MTH expects a correlation between a verb&#8217;s <italic>preference</italic> for one-patient situations and <italic>acceptance</italic> of <italic>S3</italic> situations as possible for reciprocal sentences with that verb. Testing this hypothesized correlation was the main goal of Experiment 2.</p>
<p>Part 1 of Experiment 2 was a preference task similar to the pretest of Experiment 1, but included verb concepts that were expected to show a wider range of patient cardinality preferences. This made it possible to test the MTH.<xref ref-type="fn" rid="n15">15</xref> In order to study different kinds of verbs within one experiment, both parts of Experiment 2 used a different visual presentation of the stimuli than in Experiment 1: schematic presentation of situations as in Figure <xref ref-type="fig" rid="F1">1</xref> and Figure <xref ref-type="fig" rid="F5">5</xref> (below).</p>
<fig id="F5">
<label>Figure 5</label>
<caption><p>Examples of pairs for <italic>pinch</italic> in the pretest and in part 1 of Experiment 2 (translated from Dutch).</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66647/"/>
</fig>
<p>Part 2 of Experiment 2 made use of a truth-value judgement task for testing interpretations of reciprocal sentences. The MTH makes predictions about acceptability of reciprocal sentences in situations like <italic>S3</italic> and <italic>S6</italic>. Unlike the SMH, the MTH expects there to be reciprocal sentences that are acceptable in both <italic>S3</italic> and <italic>S6</italic>. Thus, while the considerable preference rates for <italic>S3</italic> in the preference task of Experiment 1 are evidence against the SMH, we need truth-value judgements in order to support the MTH. As the strongest test for the MTH, we performed a correlation analysis to test the fine-grained relationships that the MTH predicts, between the acceptability of <italic>S3</italic> in part 2 and the preference for one-patient situations in part 1.</p>
<p>For convenience of presentation, we categorize the verbs used in Experiment 2 into three types, based on their expected patient cardinality preferences:</p>
<p><italic>Type 1 &#8211; Neutral verbs</italic>: Verbs for which we expected no preference between instances with different patient cardinality (e.g. <italic>know</italic>).</p>
<p><italic>Type 2 &#8211; One-patient-preference verbs</italic>: Verbs for which we expected a preference for situations with one patient, even though we expected two-patient situations to also be categorized as instances of the verb concept (e.g. <italic>pinch</italic>). This group of verbs was in the focus of Experiment 1.</p>
<p><italic>Type 3 &#8211; Strong one-patient-preference verbs</italic>: Verbs for which we expected a high preference for one-patient situations, because we expected two-patient situations to not be categorized as instances of the verb concept at all (e.g. <italic>bite</italic>).<xref ref-type="fn" rid="n16">16</xref></p>
<p>This intuitive classification was used in the experiment as a means for selecting candidate verbs, as well as for presenting large-scale differences between groups of verbs.</p>
<sec>
<title>4.1 Pretest</title>
<p>In preparation for Experiment 2, we conducted a pretest in Dutch that measured only patient cardinality preferences. This pretest had several aims. Firstly, it tested whether measuring typicality effects with schematic stimuli is comparable to measuring them with pictorial stimuli as in Experiment 1. Secondly, the pretest tested whether Dutch verbs behave in a comparable way to their Hebrew counterparts in Experiment 1. Thirdly, and most importantly, the results of the pretest were used for selecting the verbs to be tested in the main parts of Experiment 2. In part 1 of the experiment, we aimed to obtain a wide range of cardinality preference values. This would ensure that the eventual correlation analysis with reciprocal interpretations is based on verbs that show diverse typicality effects. To this end, the pretest was used for selecting those verbs that showed the clearest differences in patient cardinality preferences.</p>
<p>The pretest measured patient cardinality preferences for 60 Dutch verbs, 32 of which overlapped with the verbs that were used in the pretest for Experiment 1. As in Experiment 1, we measured whether there was a preference between situations with one patient per agent vs. situations with two patients per agent.</p>
<sec>
<title>4.1.1 Participants</title>
<p>Twenty-one Utrecht University students (19 female, age <italic>M</italic> = 21) participated for monetary compensation. All participants were native Dutch speakers without dyslexia. Prior to the experiment, all participants signed an informed consent.</p>
</sec>
<sec>
<title>4.1.2 Materials</title>
<p>Sixty different Dutch verbs were tested. These included 16 verbs that were a priori assumed to be of type 1 (neutral verbs like <italic>kennen</italic>, &#8216;know&#8217;), 29 verbs that were a priori classified as type 2 (one-patient-preference verbs like <italic>knijpen</italic>, &#8216;pinch&#8217;), and 8 verbs that were a priori classified as type 3 (strong one-patient-preference verbs like <italic>bijten</italic>, &#8216;bite&#8217;). In addition, we included 7 verbs like <italic>fotograferen</italic> (&#8216;photograph&#8217;) and <italic>ontmoeten</italic> (&#8216;meet&#8217;) (see Table B1 in Appendix B) that were used as filler verbs. These verbs have a typical &#8216;collective&#8217; interpretation in which an activity is performed on a collection of patients rather than on individual patients. These verbs were expected to show some preference for two patients over one. They were added to achieve an optimal balance in typicality preferences.</p>
<p>For each verb, we included one experimental pair of schematic representations, reflecting the choice between two instances of that verb &#8211; one with one patient and one with two patients. Each schema included three individuals, which were represented by three proper names, and one or more arrows between them, reflecting either one or two actions between the three individuals. An example of an experimental pair appears in the top row of Figure <xref ref-type="fig" rid="F5">5</xref>.</p>
<p>In addition, for each verb we included five filler pairs in order to control for visual complexity, by adding all possible arrow combinations, and response bias, by alternating the placement of the different configurations. The filler items differed from the experimental pair in the number of arrows (Figure <xref ref-type="fig" rid="F5">5</xref>).</p>
<p>The experimental pairs (1 per verb) and filler pairs (5 per verb) for the 60 verbs resulted in a total of 360 trials. The trials were presented in a pseudo-random order with the restriction that no verb or schematic pair would repeat in two consecutive trials. The position of the different schematic representations (on the left or on the right) and the pointing direction of the arrows (to the left or to the right) were counterbalanced over the trials and the verbs.</p>
</sec>
<sec>
<title>4.1.3 Procedure</title>
<p>The task was presented in Dutch in a sound-proof booth on a PC using <italic>Presentation</italic> software (Neurobehavioral Systems, Albany, CA). Prior to entering the booth, each participant was instructed verbally about the set-up and on how to interpret the schematic representations. Further instructions were given on the PC monitor, stressing that each schematic representation should be interpreted as representing a situation at one point in time, rather than multiple situations over a time interval. This was clarified in order to make sure that the verbs are interpreted in the same tense as in part 2 of Experiment 2 (see footnote 8). After the instructions, each participant completed six practice trials. Subsequently, participants were given the opportunity to ask for further clarifications, followed by six additional practice trials. No verbs that were used in the practice trials were used in the actual pretest. The pretest itself consisted of two blocks of trials. Each trial started with a fixation cross (500 ms), followed by the presentation of a verb in the top centre of the screen and, below the verb, a pair of schematic representations: one of the six pairs from Figure <xref ref-type="fig" rid="F5">5</xref>. Participants were instructed to select the schematic representation that best represented the given verb by pressing the left or right arrow key accordingly, with their dominant hand. The verb and the schematic representations remained visible on the screen for 5000 ms, or until the participant responded.</p>
</sec>
<sec>
<title>4.1.4 Analysis</title>
<p>We calculated the proportion of reactions to the test items where a schematic representation of a one-patient situation was selected. We performed a correlation analysis on the patient cardinality data on the 32 Hebrew verbs that were tested in the pretest for Experiment 1, and the corresponding Dutch verbs that were examined in the current pretest.</p>
</sec>
<sec>
<title>4.1.5 Results</title>
<p>The results of the pretest are given in Table B1 of Appendix B. These results about Dutch verbs show a significant one-sided positive correlation with the patient cardinality preferences of the corresponding 32 Hebrew verbs that were tested in the pretest for Experiment 1 (<italic>r</italic> (32) = .39, <italic>p</italic> (one-tailed) = .014), which indicates that the schematic method yields comparable patient cardinality preferences to the pictorial method of Experiment 1. This lends support to our assumption that the two methods measure the same feature, and that the Hebrew and Dutch verbs that we tested have similar meanings.</p>
<p>The results of the pretest were used to select verbs for Experiment 2. The selection procedure is explained in the Materials section of section 4.2.</p>
</sec>
</sec>
<sec>
<title>4.2 Experiment 2 (part 1): Typicality</title>
<p>Part 1 of Experiment 2 is a replication study that measured patient cardinality preferences for a subset of the Dutch verbs that were used in the pretest. This part again measured patient cardinality preferences for the selected subset of verbs. Therefore, we expected to see similar behavior to what we observed in the pretest: verbs that we classified as type 1 verbs were expected to show no preference, those that we classified as type 2 verbs were expected to show a preference for one-patient situations, and those that we classified as type 3 verbs were expected to show the same preference as type 2 verbs, possibly more substantially.</p>
<sec>
<title>4.2.1 Participants</title>
<p>Eighteen Utrecht University students (15 female, age <italic>M</italic> = 21) participated for monetary compensation. All participants were native Dutch speakers without dyslexia and did not participate in the pretest. Prior to the experiment, all participants signed an informed consent.</p>
</sec>
<sec>
<title>4.2.2 Materials</title>
<p>Based on the results of the pretest, we selected 18 verbs that were to be tested regarding their patient cardinality preferences (see Table B2 in Appendix B). The selection process went as follows. Firstly, to minimize confounds coming of syntactic processing we only selected transitive verbs, and ruled out verbs that would require a preposition when used in a reciprocal sentence (e.g. <italic>luisteren naar</italic> &#8216;listen to&#8217;). These verbs are marked with an asterisk in Table B1 in Appendix B. Secondly, we selected six verbs from each of the three classes of verbs. Additionally, from the type 1 verbs, we also ruled out verbs where we expected <italic>S6</italic> situations to be infelicitous for reasons that are independent of patient typicality. For example, with the verb <italic>horen</italic> (&#8216;hear&#8217;), participants showed almost an equal preference (.524) for one patient and for two patients. However, in an <italic>S6</italic> situation, a person simultaneously hears two other people while they are hearing her. Such a situation might be too noisy for the sentence <italic>they are hearing each other</italic> to make sense. Other verbs that were ruled out for similar reasons are: <italic>complimenteren</italic> (&#8216;compliment&#8217;), <italic>noemen</italic> (&#8216;name&#8217;), and <italic>zien</italic> (&#8216;see&#8217;). These verbs are marked with double asterisks in Table B1 in Appendix B. After applying this restriction, we selected the six verbs that had the lowest preference for a one-patient situation. From the type 2 verbs (one-patient-preference) as well as the type 3 verbs (strong one-patient-preference), we expected fewer confounds than with type 1 verbs. Therefore, we simply selected the six verbs from each of the two classes that showed the highest preference for a one-patient situation (on the reasons for the distinction between type 2 and type 3, see footnote 13).</p>
<p>The 18 verbs selected for Experiment 2 are the Dutch correlates to the following verbs:</p>
<table-wrap>
<table>
<tr>
<td align="left"><bold>Type 1 (neutral):</bold></td>
<td align="left"><italic>envy, know, understand, admire, miss, hate</italic></td>
</tr>
<tr>
<td align="left"><bold>Type 2 (one-patient-preference):</bold></td>
<td align="left"><italic>pinch, hit, caress, stab, shoot, grab</italic></td>
</tr>
<tr>
<td align="left"><bold>Type 3 (strong one-patient-preference):</bold></td>
<td align="left"><italic>kiss, dress, kick, lash out, bite, lick</italic></td>
</tr>
</table>
</table-wrap>
<p>Similarly to the pretest, for each of the selected 18 verbs we used one experimental pair and five filler pairs to control for visual complexity, by adding all possible arrow combinations, and response bias, by alternating the placement of the different configurations (Figure <xref ref-type="fig" rid="F5">5</xref>). This resulted in a total of 108 trial items, which were presented in a pseudo-random order with the restriction that no verb or schematic pair would repeat in two consecutive trials. The position of the different schematic representations (on the left or on the right) and the pointing direction of the arrows (to the left or to the right) were counterbalanced over the trials and the verbs.</p>
</sec>
<sec>
<title>4.2.3 Procedure</title>
<p>The procedure of the task in part 1 was identical to the procedure of the pretest.</p>
</sec>
<sec>
<title>4.2.4 Analysis</title>
<p>We calculated the proportion of reactions to the test items where a schematic representation of a one-patient situation was selected. We then performed a correlation analysis on this patient cardinality data and the patient cardinality data from the same verbs in the pretest, in order to verify that the results were comparable. Next, we performed a generalized linear mixed model (GLMM) logistic regression analysis in SPSS (v.24) on the responses to the test items (<xref ref-type="bibr" rid="B16">Quen&#233; and van den Bergh 2008</xref>). The responses to the one-patient situation were coded as 1 and to the two-patient situation were coded as 0. The model was built iteratively, starting from a basic model and adding relevant predictors one by one and assessing whether the model accuracy improved. In the final model the fixed part contained Verb type with three levels (1, 2 and 3). The random part of the model contained participants as random effect. The random effect of verb was also tested, but did not add significantly to the model, see Appendix C1 for details on model comparisons.</p>
</sec>
<sec>
<title>4.2.5 Results</title>
<p>We found a significant one-sided correlation (<italic>r</italic> (18) =.57, <italic>p</italic> (one-tailed) = .007) between the results for the 18 verbs in part 1 of Experiment 2 and those same 18 verbs in the pretest. This means that as expected, we replicated the pretest for the 18 selected verbs. The results of the one-patient preference for the 18 verbs are presented in Table <xref ref-type="table" rid="T2">2</xref>.</p>
<table-wrap id="T2">
<label>Table 2</label>
<caption><p>Experiment 2 part 1 results of one-patient preference for the tested verbs.</p></caption>
<table>
<tr>
<th align="left" style="background-color:#f3f4f4;">Verb</th>
<th align="center" style="background-color:#f3f4f4;">Type</th>
<th align="center" style="background-color:#f3f4f4;">Preference one-patient</th>
</tr>
<tr>
<td align="left" colspan="3"><hr/></td>
</tr>
<tr>
<td align="left">Grab</td>
<td align="right">1</td>
<td align="right">0.94</td>
</tr>
<tr>
<td align="left">Shoot</td>
<td align="right">1</td>
<td align="right">0.89</td>
</tr>
<tr>
<td align="left">Stab</td>
<td align="right">1</td>
<td align="right">0.88</td>
</tr>
<tr>
<td align="left">Caress</td>
<td align="right">1</td>
<td align="right">0.83</td>
</tr>
<tr>
<td align="left">Hit</td>
<td align="right">1</td>
<td align="right">0.78</td>
</tr>
<tr>
<td align="left">Pinch</td>
<td align="right">1</td>
<td align="right">0.67</td>
</tr>
<tr>
<td align="left">Lick</td>
<td align="right">2</td>
<td align="right">1.00</td>
</tr>
<tr>
<td align="left">Bite</td>
<td align="right">2</td>
<td align="right">1.00</td>
</tr>
<tr>
<td align="left">Lash out</td>
<td align="right">2</td>
<td align="right">0.94</td>
</tr>
<tr>
<td align="left">Kick</td>
<td align="right">2</td>
<td align="right">0.89</td>
</tr>
<tr>
<td align="left">Dress</td>
<td align="right">2</td>
<td align="right">0.89</td>
</tr>
<tr>
<td align="left">Kiss</td>
<td align="right">2</td>
<td align="right">0.65</td>
</tr>
<tr>
<td align="left">Hate</td>
<td align="right">3</td>
<td align="right">0.67</td>
</tr>
<tr>
<td align="left">Miss</td>
<td align="right">3</td>
<td align="right">0.44</td>
</tr>
<tr>
<td align="left">Admire</td>
<td align="right">3</td>
<td align="right">0.39</td>
</tr>
<tr>
<td align="left">Understand</td>
<td align="right">3</td>
<td align="right">0.39</td>
</tr>
<tr>
<td align="left">Know</td>
<td align="right">3</td>
<td align="right">0.33</td>
</tr>
<tr>
<td align="left">Envy</td>
<td align="right">3</td>
<td align="right">0.33</td>
</tr>
</table>
</table-wrap>
<p>For the 18 verbs in part 1 of Experiment 2, we tested the effect of Verb type (1, 2 and 3). The GLMM revealed a significant main effect of Verb type (<italic>F</italic> (2, 319) = 31.70, <italic>p</italic> &lt; .001). As expected, the results show strong preference for a one-patient situation with both type 2 verbs (<italic>M</italic> = .87, <italic>SE</italic> = .05) and type 3 verbs (<italic>M</italic> = .93, <italic>SE</italic> = .03). Pairwise comparisons showed that there was no significant difference between these preferences (<italic>t</italic>(319) = 1.38, <italic>p</italic> = .167). The preference for one-patient situations with type 1 verbs (<italic>M</italic> = .41, <italic>SE</italic> = .09) was significantly lower compared to type 2 verbs (<italic>t</italic>(319) = 6.32, <italic>p</italic> &lt; .001) and type 3 verbs (<italic>t</italic>(319) = 6.67, <italic>p</italic> &lt; .001). See Figure <xref ref-type="fig" rid="F6">6</xref>, and the further details in Table B2 in Appendix B.</p>
<fig id="F6">
<label>Figure 6</label>
<caption><p>Experiment 2 part 1 &#8211; analysis of preference for a one-patient situation with three types of verbs. Error bars represent standard errors of the mean.</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66648/"/>
</fig>
</sec>
</sec>
<sec>
<title>4.3 Experiment 2 (part 2): Reciprocal interpretation</title>
<p>Part 2 of Experiment 2 tested responses to a truth-value judgement task on Dutch reciprocal sentences in different situations. The verbs that were tested were the same verbs that were used in part 1, and the reciprocal situations were presented using schematic representations. The truth-value judgement task was used to measure to what extent a given reciprocal sentence is accepted in situations <italic>S6</italic> and <italic>S3</italic>. The MTH links typicality preferences of verbs to truth-value judgements on reciprocal sentences as follows: <italic>S6</italic> is expected to be the core situation for sentences with verbs that show no patient cardinality preferences or a preference for two-patient situations; <italic>S3</italic> is expected to be the core situation for sentences with verbs that show a preference for one-patient situations. Accordingly, we expected to see substantially higher acceptance rates with <italic>S3</italic> for verbs of type 2 and 3, compared to the same situation with verbs of type 1.</p>
<sec>
<title>4.3.1 Participants</title>
<p>A total of 25 Utrecht University students participated for monetary compensation (24 female, age <italic>M</italic> = 23). All participants were native Dutch speakers without dyslexia and did not participate in the the pretest or part 1 of the experiment. Prior to the experiment all participants signed an informed consent.</p>
</sec>
<sec>
<title>4.3.2 Materials</title>
<p>The same 18 verbs from part 1 of Experiment 2 were used, but now in Dutch reciprocal sentences of the form <italic>A, B and C P each other</italic> (where <italic>A, B</italic> and <italic>C</italic> are proper names and <italic>P</italic> is a verb). The resulting 18 sentences were tested for their interpretation in a truth-value judgement task.</p>
<p>For each sentence, we included two experimental trials &#8211; schemas reflecting <italic>S6</italic> and <italic>S3</italic>. As in Experiment 1, those situations illustrate relations between three individuals. In situation <italic>S6</italic>, each individual acts on both other individuals. In situation <italic>S3</italic>, each individual acts on exactly one other individual and is acted on by exactly one other individual. Each schema included three individuals, which were represented by three proper names, and either six arrows (in <italic>S6</italic>) or three arrows (in <italic>S3</italic>) between them. Examples of these experimental trials are in the top two rows of Figure <xref ref-type="fig" rid="F7">7</xref>.</p>
<fig id="F7">
<label>Figure 7</label>
<caption><p>Examples of trials for <italic>John, Bill and George pinch each other</italic> in part 2 of Experiment 2 (translated from Dutch).</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66649/"/>
</fig>
<p>In addition, for each verb we included one control trial and two filler trials (Figure <xref ref-type="fig" rid="F7">7</xref>). The filler trials were used to control for visual complexity and frequency of no-responses. The control trials illustrated a situation in which only two individuals act on another individual (<italic>S2</italic>). We added this control item because it was expected to be a typical instance for all verbs, but not the core situation for any reciprocal sentence containing them. Similarly, for sentences with verbs of type 1, situation <italic>S3</italic> is expected to be typical but not the core situation. By adding the control trial <italic>S2</italic>, we aimed to also have typical but non-core situations for sentences with verbs of types 2 and 3. This allowed us to better test the predictions of the proposed MTH, since this principle also makes predictions about acceptability rates for non-core situations.</p>
<p>The experimental trials (2 per verb), control trial (1 per verb) and filler trials (2 per verb) for all 18 sentences resulted in a total of 90 trials. The trials were presented in a pseudo-random order with the restriction that no verb or schematic representation would repeat in two consecutive items. The pointing direction of the arrows (to the left or to the right and up or down) was counterbalanced over the verbs.</p>
</sec>
<sec>
<title>4.3.3 Procedure</title>
<p>Part 2 of Experiment 2 consisted of two blocks of trials. As in part 1, the task was presented in Dutch in a sound-proof booth on a PC with <italic>Presentation</italic> software (Neurobehavioral Systems, Albany, CA). The instructions on the way schemas were to be interpreted resembled those of part 1, and were similarly followed by practice trials.</p>
<p>Each trial started with a fixation cross (500 ms), followed by the presentation of a reciprocal sentence (e.g. <italic>John, Bill en George knijpen elkaar</italic>, &#8216;John, Bill and George pinch each other&#8217;) at the top of the screen. After 2000 ms, a schematic representation was added to the screen, below the sentence. Participants were instructed to indicate whether the situation that was presented schematically is a possible depiction of the sentence or not, by pressing a green or red button accordingly (right and left arrow key respectively, marked with a sticker), with their dominant hand. The sentence and the schematic representation remained visible on the screen until the participant responded, or for 10000 milliseconds if there was no response.</p>
</sec>
<sec>
<title>4.3.4 Analysis</title>
<p>We calculated the proportion of affirmative responses to the control trials (<italic>S2</italic>) and the experimental trials (<italic>S6</italic> and <italic>S3</italic>), which reflects the acceptability of a given schematic representation of a reciprocal situation as a possible depiction of the given sentence. Further statistical analysis focused on the experimental trials (<italic>S6</italic> and <italic>S3</italic>). We performed a generalized linear mixed model (GLMM) logistic regression analysis in SPSS (v.24) on the responses to the experimental trials. The model was built iteratively, starting from a basic model and adding relevant predictors one by one and assessing whether the model accuracy improved. In the final model the fixed part contained Verb type with three levels (1, 2 and 3) and Trial type with two levels (<italic>S6</italic> and <italic>S3</italic>). The random part of the model contained participants as random effect. The random effect of verb was also tested, but did not add significantly to the model, see Appendix C2 for details on model comparisons.</p>
</sec>
<sec>
<title>4.3.5 Results</title>
<p><bold><italic>Control trial S2</italic>:</bold> The acceptance of reciprocal sentences in <italic>S2</italic> was very low for sentences with all verb types: reciprocal sentences with type 1 verbs (<italic>M</italic> = .06, <italic>SE</italic> = .02), reciprocal sentences with type 2 verbs (<italic>M</italic> = .13, <italic>SE</italic> = .06), and reciprocal sentences with type 3 verbs (<italic>M</italic> = .12, <italic>SE</italic> = .05). For details see Table B2 in Appendix B.</p>
<p>Further analysis of the results of part 2 of Experiment 2 focused on the two experimental trial types that measured acceptability of reciprocal sentences in <italic>S6</italic> and <italic>S3</italic>. There was a significant main effect of Trial type (<italic>F</italic> (1, 893) = 10.18, <italic>p</italic> = .001) and Verb type (<italic>F</italic> (2, 893) = 11.52, <italic>p</italic> &lt; .001), as well as a significant interaction between Trial type and Verb type (<italic>F</italic> (2, 893) = 30.06, <italic>p</italic> &lt; .001). This interaction was further analyzed with pairwise comparisons.</p>
<p><bold><italic>Experimental trial S6</italic>:</bold> Regarding the acceptance of reciprocal sentences in <italic>S6</italic>, there were significant differences between all three types of sentences (all <italic>p&#8217;s</italic> &lt; .001) (see Figure <xref ref-type="fig" rid="F8">8</xref> and Table B2 in Appendix B). The highest acceptability of sentences in <italic>S6</italic> was found for those sentences containing type 1 verbs (<italic>M</italic> = .98, <italic>SE</italic> = .01), followed by those with type 2 verbs (<italic>M</italic> = .87, <italic>SE</italic> = .03) and those with type 3 verbs (<italic>M</italic> = .60, <italic>SE</italic> = .05).</p>
<fig id="F8">
<label>Figure 8</label>
<caption><p>Experiment 2 part 2 &#8211; item analysis of acceptance rate for sentences in <italic>S6</italic> and <italic>S3</italic> with 3 types of verbs. Error bars represent standard errors of the mean.</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66650/"/>
</fig>
<p><bold><italic>Experimental trial S3</italic>:</bold> Regarding acceptance of reciprocal sentences in <italic>S3</italic>, the pattern was different (see Figure <xref ref-type="fig" rid="F8">8</xref> and Table B2 in Appendix B). There was no significant difference between reciprocal sentences with type 2 verbs (<italic>M</italic> = .89, <italic>SE</italic> = .03) and reciprocal sentences with type 3 verbs (<italic>M</italic> = .84, <italic>SE</italic> = .04, <italic>t</italic> (893) = 1.32, <italic>p</italic> = .189). Both types of sentences showed high levels of acceptability in <italic>S3</italic>. The acceptability of reciprocal sentences with type 1 verbs in <italic>S3</italic> (<italic>M</italic> = .53, <italic>SE</italic> = .065) was significantly lower compared to sentences with type 2 verbs (<italic>t</italic> (893) = 6.74, <italic>p</italic> &lt; .001) and with type 3 verbs (<italic>t</italic> (893) = 5.76, <italic>p</italic> &lt; .001).</p>
</sec>
</sec>
<sec>
<title>4.4 Correlation between results of part 1 and part 2</title>
<p>As expected, the results from part 1 showed a considerable variability in the patient cardinality preferences of different verbs (<italic>M</italic> = .72, <italic>SD</italic> = .24, Table B2 in Appendix B). The MTH predicts that the preference for one-patient situations (as measured in part 1) correlates with the acceptability of reciprocal sentences in <italic>S3</italic> (as measured in part 2), explaining the interpretations of reciprocal sentences with type 2 verbs. This prediction was borne out: we found a significant one-sided positive correlation (<italic>r</italic> (18) = .76, <italic>p</italic> &lt; .001, see Figure <xref ref-type="fig" rid="F9">9</xref>).</p>
<fig id="F9">
<label>Figure 9</label>
<caption><p>Relation between preferences for one-patient situations in part 1 of Experiment 2, and acceptance of reciprocal sentences in <italic>S3</italic> in part 2 of Experiment 2.</p></caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/4986/file/66651/"/>
</fig>
</sec>
<sec>
<title>4.5 Discussion</title>
<p>The main aim of Experiment 2 was to test the predictions of the proposed Maximal Typicality Hypothesis: whether the core situation described by a reciprocal sentence is maximal among those situations that are most typical for the verb concept in the reciprocal&#8217;s scope. To that end, part 1 measured typicality effects for 18 different verb concepts, which were selected based on their patient cardinality preferences as measured in a pretest. Part 2 measured whether reciprocal sentences containing those 18 verbs were accepted in <italic>S6</italic> and/or <italic>S3</italic> (i.e. the two experimental trials). The test for the MTH was in the observed differences between verb types, and in the more fine-grained correlation analysis between the two parts of Experiment 2.</p>
<sec>
<title>4.5.1 Main results: The core situation</title>
<p>In part 1 of the experiment we found that different verbs show different patient cardinality preferences. As expected, those verbs that we a priori classified as type 1 (neutral) showed no clear preference for an instance with one patient vs. two patients. By contrast, those verbs that we called type 2 (one-patient-preference) and type 3 (strong one-patient-preference) showed a clear preference for instances with one patient over instances with two patients. From these patient cardinality preferences, we extrapolate typicality preferences for reciprocal situations. For type 1 verbs, we extrapolate that there is no difference in typicality between reciprocal situations that differ in terms of patient cardinality, while for verbs of type 2 and 3 we extrapolate that reciprocal situations with one patient per agent are more typical than those with two patients per agent. If, as extrapolated, all situations for type 1 verbs are of equal typicality, then the MTH predicts that the core situation for a reciprocal sentence with a type 1 verb is <italic>S6</italic>: the maximal situation among all situations. Thus, for reciprocal sentences with type 1 verbs the MTH expects the acceptability in <italic>S6</italic> in part 2 of the experiment to be high, and substantially higher than in <italic>S3</italic>. For sentences with type 2 and type 3 verbs, the MTH predicts the core situation to be <italic>S3</italic>: the maximal among those situations that have no more than one patient per agent. Thus, the MTH predicts the acceptability of these sentences to be high in <italic>S3</italic>, and substantially higher than the acceptability of reciprocal sentences with type 1 verbs in <italic>S3</italic> situations.</p>
<p>The results of part 2 show that reciprocal sentences with type 1 verbs are very often accepted in <italic>S6</italic> (98% of the time), and substantially more so than in <italic>S3</italic> (53%, <italic>t</italic> (893) = 8.22, <italic>p</italic> &lt; .001). Our findings also indicate that sentences with type 3 verbs are highly acceptable in <italic>S3</italic> (83%), significantly more so than sentences with type 1 verbs. These findings are expected by the MTH based on the typicality preferences observed in Part 1. The same results are also explained by the SMH, provided that we assume that type 3 verbs like <italic>bite</italic> are judged as physically impossible in situations with two patients per agent (an assumption that we did not test directly in our experiments).</p>
<p>However, when it comes to type 2 verbs, our findings offer substantial support for the MTH over the SMH. Reciprocal sentences with such verbs showed high acceptability rates in <italic>S3</italic> (88%), comparable to type 3 verbs, and significantly more so than sentences with type 1 verbs. This is as expected by the MTH, while the SR meaning that the SMH predicts for these sentences does not hold in <italic>S3</italic>.<xref ref-type="fn" rid="n17">17</xref></p>
<p>The success of the MTH in predicting the acceptability pattern for reciprocal sentences in <italic>S3</italic> is further supported by the significant correlation that was found between this acceptability and preferences for one-patient situations as measured in part 1. In this way, the MTH not only explains the overall acceptability of reciprocal sentences with type 2 verbs in <italic>S3</italic>, but gives a fine-grained prediction for the interpretation of other reciprocal sentences in <italic>S3</italic> situations.</p>
</sec>
<sec>
<title>4.5.2 Additional results: Non-core situations</title>
<p>So far, we have only discussed the acceptability of reciprocal sentences in the core situation. That is, the high acceptability of sentences with type 1 verbs in <italic>S6</italic>, and the high acceptability of sentences with type 2 and type 3 verbs in <italic>S3</italic>. In addition to that, however, our data provide information about the acceptability of these sentences in non-core situations.</p>
<p>Firstly, one important aspect of our results are the observed acceptability rates of sentences with type 2 and type 3 verbs in <italic>S6</italic>. For both kinds of sentences, the MTH predicts the same core situation: <italic>S3</italic>. However, as we might expect, the acceptability of reciprocal sentences with type 3 verbs (e.g. <italic>bite</italic>) in <italic>S6</italic> is significantly lower than that of sentences with type 2 verbs (e.g. <italic>pinch</italic>) (<italic>t</italic> (893) = &#8211;5.14, <italic>p</italic> &lt; .001). Sentences with type 2 verbs were in fact equally acceptable in <italic>S6</italic> and <italic>S3</italic> (<italic>t</italic> (893) = 0.69, <italic>p</italic> = .492). Potential differences between type 2 verbs and type 3 verbs with respect to two-patient situations like <italic>S6</italic> were not measured in the preference tasks of part 1. Thus, the difference in interpretations between sentences with type 2 verbs and type 3 verbs calls for a separate explanation. We hypothesize that the high preference for one-patient situations with such verbs (Experiment 2 part 1) has an additional factor with type 3 verbs compared to type 2 verbs. In the case of type 3 verbs, we believe that for many participants, the one-patient preference reflects not merely a choice between two possible situations, but a preference of a possible situation (with one patient) over an impossible, or inconceivable, situation (with two patients). For type 2 verbs, one-patient situations are uniformly preferred over two-patient situations, but the latter are likely to be accepted as instances of the verb concept, as witnessed by the high acceptability rates of sentences with such verbs in <italic>S6</italic>. Such a difference between type 2 and type 3 verbs could not be measured in the forced choice task of part 1, but it would immediately affect acceptability judgements of reciprocal sentences in <italic>S6</italic> in part 2. If a participant thinks that a two-patient situation is not possible for type 3 verbs like <italic>bite</italic>, she will judge <italic>S6</italic> as strictly speaking impossible for a reciprocal sentence with such verbs. For type 2 verbs, two-patient situations put less strain on the imagination of the participants. Accordingly, <italic>S6</italic> situations are usually accepted for reciprocal sentences with type 2 verbs.<xref ref-type="fn" rid="n18">18</xref> This pattern with <italic>S6</italic> situations and type 2 and 3 verbs does not come as a surprise: it fully agrees with our <italic>a priori</italic> classification of type 2 and type 3 verbs, which was based on mere introspection, and it is consistent with the results of Experiment 1 on type 2 verbs and the assumption of SMH-based accounts on type 3 verbs. Therefore, we believe that simple experimental measures can distinguish type 2 verbs from type 3 verbs: e.g. asking participants to mark possible situations, rather than to choose between them. Running such an experiment would be unproblematic, if relevant for further research.</p>
<p>Another question on non-core situations concerns the status of <italic>S3</italic> for sentences with type 1 verbs. For such verbs, <italic>S3</italic> situations are properly contained in their core situation: <italic>S6</italic>. As expected, reciprocal sentences with those verbs are fully acceptable in <italic>S6</italic>. The SMH accounts for this fact using the strong reciprocity operator, which expects these sentences to be downright unacceptable in all situations that are properly contained by <italic>S6</italic>. The acceptability rates of sentences with type 1 verbs in <italic>S3</italic> in our experiment (<italic>M</italic> = .60, <italic>SE</italic> = .06) go against this prediction of the SMH, especially when compared to the downright unacceptability of the same sentences in <italic>S2</italic>. This problem for the SMH appears because it makes standard binary (true/false) predictions about reciprocal sentences. Accordingly, the SMH expects decisive judgements on sentences in all situations. Unlike the SMH, the MTH does not use such absolute terms for acceptability of reciprocal sentences in situations that are properly contained in the core situation. The MTH only expects this acceptability to be lower than the acceptability in the core situation. Thus, sentences with type 1 verbs are expected to show decreasing acceptability in <italic>S2</italic> and <italic>S3</italic> situations compared to <italic>S6</italic>. Similarly, sentences with type 2 and 3 verbs are expected to show lower acceptability in <italic>S2</italic> than in <italic>S3</italic>. Our acceptability results on <italic>S2</italic> and <italic>S3</italic> are consistent with these predictions.</p>
<p>The acceptability rates in <italic>S3</italic> are also relevant for evaluating another influential analysis of reciprocals. According to Langendoen&#8217;s (<xref ref-type="bibr" rid="B6">1978</xref>) operator of Weak Reciprocity, a sentence like <italic>the girls know each other</italic> should mean &#8220;every girl knows another girl and is known by another girl&#8221;, which is true in the <italic>S3</italic> situation. However, only 36% of the participants in Experiment 2 accepted the Dutch sentence with the verb &#8220;know&#8221;. More generally, our results show a disadvantage for Langendoen&#8217;s account for reciprocals with type 1 verbs. Reciprocal sentences with these verbs are significantly less acceptable in <italic>S3</italic> than in <italic>S6</italic>, whereas Langendoen&#8217;s Weak Reciprocity is equally satisfied by both situations. Thus, some selection principle should be superimposed on any analysis of reciprocals, and explain the marginality of type 1 verbs in <italic>S3</italic> situations, as opposed to other verbs. This is in line with the main conclusions of the present paper.<xref ref-type="fn" rid="n19">19</xref></p>
</sec>
</sec>
</sec>
<sec>
<title>5 Conclusion</title>
<p>Previous work has predicted that sentences like <italic>the boys know each other</italic> and <italic>the boys pinch each other</italic> are both true only if each of the boys knows/pinches all of the other boys. Our experiments show that this reading is indeed preferred for verbs like <italic>know</italic>. However, with verbs like <italic>pinch</italic>, reciprocals show a weaker meaning. This challenge for previous accounts is addressed by considering that situations where one boy pinches two boys simultaneously are quite atypical. According to our proposed account, when speakers observe this kind of non-typicality, it boosts weaker interpretations of reciprocal sentences. This tendency explains why sentences like <italic>John, Bill and George pinch each other</italic> are as acceptable when each boy pinches only one other boy, as when each boy (atypically) pinches two other boys. Furthermore, our experiments showed that the more atypical the interpretation of a reciprocal sentence, the higher the tendency of speakers to accept weaker interpretations. These findings were formally described using a new principle, the <italic>Maximal Typicality Hypothesis</italic>, which specifies a meaning for a reciprocal expression based on the lexical preferences of the predicate it combines with.</p>
</sec>
<sec sec-type="supplementary-material">
<title>Additional files</title>
<p>The additional files for this article can be found as follows:</p>
<supplementary-material id="S1" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/gjgl.180.s1">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="gjgl-3-180-s1.docx">gjgl-3-180-s1.docx</inline-supplementary-material>]-->
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="gjgl-3-180-s1.pdf">gjgl-3-180-s1.pdf</inline-supplementary-material>]-->
<label>Appendix A</label>
<caption><p>Results Experiment 1. DOI: <uri>https://doi.org/10.5334/gjgl.180.s1</uri></p></caption>
</supplementary-material>
<supplementary-material id="S2" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/gjgl.180.s1">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="gjgl-3-180-s1.docx">gjgl-3-180-s1.docx</inline-supplementary-material>]-->
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="gjgl-3-180-s1.pdf">gjgl-3-180-s1.pdf</inline-supplementary-material>]-->
<label>Appendix B</label>
<caption><p>Results Experiment 2. DOI: <uri>https://doi.org/10.5334/gjgl.180.s1</uri></p></caption>
</supplementary-material>
</sec>
</body>
<back>
<sec>
<title>Abbreviations</title>
<p><sc>ACC</sc> = accusative; <sc>MASC</sc> = masculine; <sc>PRES</sc> = present; <sc>SG</sc> = singular.</p>
</sec>
<fn-group>
<fn id="n1"><p>Throughout this paper, we use the term &#8220;relation&#8221; informally, to refer to ordered pairs of individuals. Sets of such pairs (=standard <italic>binary relations</italic>) are informally referred to as &#8220;situations&#8221;. This non-standard terminology is convenient for presentational purposes.</p></fn>
<fn id="n2"><p>Further work on reciprocals in Murray (<xref ref-type="bibr" rid="B10">2007</xref>; <xref ref-type="bibr" rid="B11">2008</xref>) and Dotla&#269;il (<xref ref-type="bibr" rid="B3">2013</xref>) discusses their behavior in discourse and suggests theories that substantially diverge from the standard proposals of Dalrymple et al. and Heim et al. that take reciprocals to be standard quantificational operators, possibly with an additional anaphoric element. Murray and Dotla&#269;il&#8217;s works do not systematically address the semantic/pragmatic factors that affect the choice of a specific reciprocal interpretation, and the issues they deal with are largely orthogonal to the main topic of the present work. The present paper develops the work in Kerem et al. (<xref ref-type="bibr" rid="B5">2009</xref>), which has also been followed in other areas: see Poortman&#8217;s (<xref ref-type="bibr" rid="B12">2014</xref>; <xref ref-type="bibr" rid="B13">2017a</xref>; <xref ref-type="bibr" rid="B14">b</xref>) experimental finding that the interpretation of plural predicate conjunctions as in <italic>the shapes are small and big/blue and big</italic> is governed by a principle that follows the Maximal Typicality Hypothesis proposed in the present paper. For other recent experimental and theoretical work which is relevant to Poortman&#8217;s findings, see Poortman &amp; Pylkk&#228;nen (<xref ref-type="bibr" rid="B15">2016</xref>), Lee (<xref ref-type="bibr" rid="B7">2017</xref>), Scontras &amp; Goodman (<xref ref-type="bibr" rid="B21">2017</xref>) and Winter (<xref ref-type="bibr" rid="B23">2017</xref>).</p></fn>
<fn id="n3"><p>For precision, we should note that <italic>S6</italic> actually represents a class of situations where each agent acts on both other agents, and possibly itself. Since this last formal detail is irrelevant for our discussion, we sloppily refer to <italic>S6</italic> as one situation.</p></fn>
<fn id="n4"><p>For consistency with Dalrymple et al.&#8217;s terminology, we refer here to the kind of world knowledge about predicates like <italic>know</italic> and <italic>bite</italic> as being part of the &#8220;context&#8221;. Whether this term is adequate depends on one&#8217;s conception of lexical meaning and pragmatic context, but it is irrelevant for our purposes.</p></fn>
<fn id="n5"><p>The MTH was first proposed in Kerem et al. (<xref ref-type="bibr" rid="B5">2009</xref>). The present work develops this idea and tests it more extensively.</p></fn>
<fn id="n6"><p>Our use of the term &#8220;interpretation&#8221; is meant to highlight our conviction that the semantic truth-conditional analysis of plural sentences, including reciprocal sentences, should be relativized to speakers&#8217; beliefs about the predicate concepts in such sentences, as well as other contextual parameters.</p></fn>
<fn id="n7"><p>Note that in this analysis, there may be two or more situations that attain such a maximum, but without any of these situations containing the other. In such cases, the MTH would define more than one situation as a core situation. Such cases do not surface in our experiments. For this reason, here and henceforth we assume uniqueness of the core situation, and refer to it as &#8216;<bold>the</bold> core situation&#8217;.</p></fn>
<fn id="n8"><p>The verbs like <italic>pinch</italic> and <italic>know</italic> categorize events and states, respectively. The event/state distinction affects typicality preferences with verbs, hence the predictions of the MTH. However, since it does not directly affect the formulation of the MTH, we will not focus on it here. A similar point holds for tense: event verbs like <italic>pinch</italic> can take the progressive tense, while state verbs like <italic>know</italic> cannot. The choice of tense naturally affects typicality preferences. However, as we shall see, the MTH is about the correlation between typicality preferences and reciprocal interpretation. For this reason, holding the tense variable constant in the two measures does not affect the predicted correlation between them.</p></fn>
<fn id="n9"><p>We here ignore situations that neither contain nor are contained by a core situation. As remarked in footnote 7, what we call <italic>the</italic> core situation should more accurately be called the <italic>class</italic> of maximal situations that attain maximal typicality. By definition, every other situation either contains or is contained by a situation in this class. For instance, <italic>S3</italic> situations were described as the situations where each of three participants only acts on one other participant. Assuming that participants do not act on themselves, there are six situations like that. Now, each of the situations where participants do not act on themselves either contains or is contained by one of these six core situations.</p></fn>
<fn id="n10"><p>Note that the MTH does not make any prediction about the rate in which acceptability declines. Thus, in principle, <italic>S2</italic> situations may be as (un)acceptable for (2)/(3) as they are for (1), despite the different core situations of these sentences (<italic>S3</italic> vs. <italic>S6</italic>, respectively). The question of how strongly acceptability declines in situations that are contained in the core situation is left for further research.</p></fn>
<fn id="n11"><p>An anonymous <italic>Glossa</italic> reviewer points out that this consideration about felicity of <italic>S6</italic> situations may give the impression that the appeal to typicality in our MTH proposal is dispensable. However, such an impression would be misleading. The MTH is a hypothesis about truth-conditions, based on typicality information. On top of any such theory of truth, we should always consider the plausibility of models that support sentences as true. For instance, the non-reciprocal sentences &#8220;Mary is pinching (biting) Dan and Max simultaneously&#8221; are expected to be true if Mary is pinching (biting) Dan while pinching (biting) Max at the same time. Due to the marginal status of the &#8220;simultaneous bite&#8221; scenario, the corresponding sentence is unacceptable. This kind of consideration is fairly standard, and is orthogonal to both the semantics of reciprocity and the phenomenon of typicality preferences.</p></fn>
<fn id="n12"><p>In fact, the latter situation may even be considered physically impossible. This means that situations with multiple patients per agent may not be instances of the verb concept <italic>bite</italic>.</p></fn>
<fn id="n13"><p>This phrasing is not perfect from a theoretical semantic point of view, but it proved clearest for the participants.</p></fn>
<fn id="n14"><p>An anonymous reviewer points out that illustrations might be discarded as physically unrealistic. If that were the case in Experiment 1, it might in principle salvage the SMH, since those participants who preferred <italic>S3</italic> might have done so because they considered the illustration representing <italic>S6</italic> as impossible. However, if that was the case, we would expect that in a truth-value judgement task with the same verbs, participants would not accept the corresponding reciprocal sentences in <italic>S6</italic>. As we will see, in Experiment 2 the vast majority of participants did accept the parallel reciprocal sentences in Dutch in <italic>S6</italic> situations. Thus, we conclude that the possibility that the reviewer suggests was also not the case with the <italic>S3</italic> preference data in Experiment 1.</p></fn>
<fn id="n15"><p>This is in contrast to Experiment 1, where the goal was to test the SMH using concepts with a one-patient preference. In Experiment 1, this goal dictated a selection of verbs that showed low distribution of such preferences (<italic>M</italic> = .79, <italic>SD</italic> = .08 for the 11 target verbs that were selected for Experiment 1). The two control verbs (<italic>hug</italic> and <italic>give a speech</italic>), which were intended to add more variation, instead showed unexpected effects (see Table <xref ref-type="table" rid="T1">1</xref>). This is likely due to factors other than patient cardinality, for example the difficulty of illustrating those verbs in <italic>S3</italic> and <italic>S6</italic>, and the influence of other typicality parameters such as preference for one <italic>agent</italic> acting on any patient (rather than more agents).</p></fn>
<fn id="n16"><p>Note that the difference between type 2 and type 3 verbs is not expected to surface when measuring the preference for one-patient situations over two-patient situations (as in part 1 of Experiment 2). However, the difference between the two types is expected to show in truth-value judgments of reciprocal sentences in situations that contain two-patient situations, i.e. <italic>S6</italic> (in part 2 of Experiment 2). Therefore, the two groups are distinguished here, although the difference between them is not measured in the preference task of Part 1.</p></fn>
<fn id="n17"><p>Note that with type 2 verbs, the acceptability rates of reciprocal sentences in <italic>S6</italic> situations were as high as in <italic>S3</italic> situations. This means that SR cannot be ruled out, and accordingly the SMH expects <italic>S3</italic> not to be accepted (see also footnote 14 above).</p></fn>
<fn id="n18"><p>This is consistent with the results of the forced-choice task of Experiment 1, where <italic>S6</italic> was often preferred to <italic>S3</italic> for type 2 verbs, despite the high typicality of one-patient situations.</p></fn>
<fn id="n19"><p>A similar point is made by Dalrymple et al. (<xref ref-type="bibr" rid="B2">1998: 165</xref>), on the basis of their intuitive judgements about the sentence &#8220;House of Commons etiquette requires legislators to address only the speaker of the House and <italic>refer to each other indirectly</italic>&#8221;.</p></fn>
</fn-group>
<ack>
<title>Acknowledgements</title>
<p>Eva B. Poortman and Marijn E. Struiksma1 contributed equally to this paper.</p>
<p>The work by the first, second and fifth author was partially supported by a VICI grant 277-80-002 of the Netherlands Organisation for Scientific Research (NWO). Work by the first and last author was also partially funded by the European Research Council (ERC) under the European Union&#8217;s Horizon 2020 research and innovation programme (grant agreement No 742204). Work by the fifth author was also partially supported by an NWO grant &#8220;Reciprocal Expressions and Relational Processes in Language&#8221; (2007/8). The work by the third author was partially supported by the Israel Science Foundation (grant 2005231). Figures <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F3">3</xref>, <xref ref-type="fig" rid="F4">4</xref> and all drawings used in Experiment 1 are by Ruth Noy Shapira. We are grateful to Michal Biran for help with Experiment 1, and to Gideon Keren and Alda Mari for extensive comments on a previous draft of this paper.</p>
</ack>
<sec>
<title>Competing Interests</title>
<p>The authors have no competing interests to declare.</p>
</sec>
<ref-list>
<ref id="B1"><label>1</label><mixed-citation publication-type="journal"><string-name><surname>Beck</surname>, <given-names>Sigrid</given-names></string-name>. <year>2001</year>. <article-title>Reciprocals are definites</article-title>. <source>Natural Language Semantics</source> <volume>9</volume>. <fpage>69</fpage>&#8211;<lpage>138</lpage>. DOI: <pub-id pub-id-type="doi">10.1023/A:1012203407127</pub-id></mixed-citation></ref>
<ref id="B2"><label>2</label><mixed-citation publication-type="journal"><string-name><surname>Dalrymple</surname>, <given-names>Mary</given-names></string-name>, <string-name><given-names>Makoto</given-names> <surname>Kanazawa</surname></string-name>, <string-name><given-names>Yookyung</given-names> <surname>Kim</surname></string-name>, <string-name><given-names>Sam</given-names> <surname>Mchombo</surname></string-name> &amp; <string-name><given-names>Stanley</given-names> <surname>Peters</surname></string-name>. <year>1998</year>. <article-title>Reciprocal expressions and the concept of reciprocity</article-title>. <source>Linguistics and Philosophy</source> <volume>21</volume>. <fpage>159</fpage>&#8211;<lpage>210</lpage>. DOI: <pub-id pub-id-type="doi">10.1023/A:1005330227480</pub-id></mixed-citation></ref>
<ref id="B3"><label>3</label><mixed-citation publication-type="journal"><string-name><surname>Dotla&#269;il</surname>, <given-names>Jakub</given-names></string-name>. <year>2013</year>. <article-title>Reciprocals distribute over information states</article-title>. <source>Journal of Semantics</source> <volume>30</volume>(<issue>4</issue>). <fpage>423</fpage>&#8211;<lpage>477</lpage>. DOI: <pub-id pub-id-type="doi">10.1093/jos/ffs016</pub-id></mixed-citation></ref>
<ref id="B4"><label>4</label><mixed-citation publication-type="journal"><string-name><surname>Heim</surname>, <given-names>Irene</given-names></string-name>, <string-name><given-names>Howard</given-names> <surname>Lasnik</surname></string-name> &amp; <string-name><given-names>Robert</given-names> <surname>May</surname></string-name>. <year>1991</year>. <article-title>Reciprocity and plurality</article-title>. <source>Linguistic Inquiry</source> <volume>22</volume>. <fpage>63</fpage>&#8211;<lpage>101</lpage>.</mixed-citation></ref>
<ref id="B5"><label>5</label><mixed-citation publication-type="journal"><string-name><surname>Kerem</surname>, <given-names>Nir</given-names></string-name>, <string-name><given-names>Naama</given-names> <surname>Friedmann</surname></string-name> &amp; <string-name><given-names>Yoad</given-names> <surname>Winter</surname></string-name>. <year>2009</year>. <article-title>Typicality effects and the logic of reciprocity</article-title>. In <string-name><given-names>Ed</given-names> <surname>Cormany</surname></string-name>, <string-name><given-names>Satoshi</given-names> <surname>Ito</surname></string-name> &amp; <string-name><given-names>David</given-names> <surname>Lutz</surname></string-name> (eds.), <source>Proceedings of Semantics and Linguistic Theory (SALT)</source> <volume>19</volume>. <fpage>257</fpage>&#8211;<lpage>274</lpage>. eLanguage. DOI: <pub-id pub-id-type="doi">10.3765/salt.v19i0.2537</pub-id></mixed-citation></ref>
<ref id="B6"><label>6</label><mixed-citation publication-type="journal"><string-name><surname>Langendoen</surname>, <given-names>D. Terence</given-names></string-name>. <year>1978</year>. <article-title>The logic of reciprocity</article-title>. <source>Linguistic Inquiry</source> <volume>9</volume>(<issue>2</issue>). <fpage>177</fpage>&#8211;<lpage>197</lpage>.</mixed-citation></ref>
<ref id="B7"><label>7</label><mixed-citation publication-type="book"><string-name><surname>Lee</surname>, <given-names>Choonkyu</given-names></string-name>. <year>2017</year>. <chapter-title>Typicality knowledge and the interpretation of adjectives</chapter-title>. In <string-name><given-names>James A.</given-names> <surname>Hampton</surname></string-name> &amp; <string-name><given-names>Yoad</given-names> <surname>Winter</surname></string-name> (eds.), <source>Compositionality and Concepts in Linguistics and Psychology</source> (Language, Cognition, and Mind) <volume>3</volume>. <fpage>123</fpage>&#8211;<lpage>138</lpage>. <publisher-name>Springer</publisher-name>, <publisher-loc>Cham</publisher-loc>. DOI: <pub-id pub-id-type="doi">10.1007/978-3-319-45977-6_5</pub-id></mixed-citation></ref>
<ref id="B8"><label>8</label><mixed-citation publication-type="book"><string-name><surname>Margolis</surname>, <given-names>Eric</given-names></string-name> &amp; <string-name><given-names>Stephen</given-names> <surname>Laurence</surname></string-name>. <year>1999</year>. <chapter-title>Concepts and cognitive science</chapter-title>. In <string-name><given-names>Eric</given-names> <surname>Margolis</surname></string-name> &amp; <string-name><given-names>Stephen</given-names> <surname>Laurence</surname></string-name> (eds.), <source>Concepts: Core readings</source>, <fpage>3</fpage>&#8211;<lpage>82</lpage>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</mixed-citation></ref>
<ref id="B9"><label>9</label><mixed-citation publication-type="journal"><string-name><surname>Mari</surname>, <given-names>Alda</given-names></string-name>. <year>2013</year>. <article-title><italic>Each Other</italic>, asymmetry and reasonable futures</article-title>. <source>Journal of Semantics</source> <volume>31</volume>(<issue>2</issue>). <fpage>209</fpage>&#8211;<lpage>261</lpage>. DOI: <pub-id pub-id-type="doi">10.1093/jos/fft003</pub-id></mixed-citation></ref>
<ref id="B10"><label>10</label><mixed-citation publication-type="confproc"><string-name><surname>Murray</surname>, <given-names>Sarah E.</given-names></string-name> <year>2007</year>. <article-title>Dynamics of reflexivity and reciprocity</article-title>. In <string-name><given-names>Maria</given-names> <surname>Aloni</surname></string-name>, <string-name><given-names>Paul</given-names> <surname>Dekker</surname></string-name> &amp; <string-name><given-names>Floris</given-names> <surname>Roelofsen</surname></string-name> (eds.), <conf-name>Proceedings of the Sixteenth Amsterdam Colloquium</conf-name>, <fpage>157</fpage>&#8211;<lpage>162</lpage>. <conf-loc>Amsterdam</conf-loc>: <conf-sponsor>Institute for Logic, Language, and Computation</conf-sponsor>.</mixed-citation></ref>
<ref id="B11"><label>11</label><mixed-citation publication-type="book"><string-name><surname>Murray</surname>, <given-names>Sarah E.</given-names></string-name> <year>2008</year>. <chapter-title>Reflexivity and reciprocity with(out) underspecification</chapter-title>. In <string-name><given-names>Atle</given-names> <surname>Gr&#248;nn</surname></string-name> (ed.), <source>Proceedings of Sinn und Bedeutung</source> <volume>12</volume>. <fpage>455</fpage>&#8211;<lpage>469</lpage>. <publisher-loc>Oslo</publisher-loc>: <publisher-name>ILOS</publisher-name>.</mixed-citation></ref>
<ref id="B12"><label>12</label><mixed-citation publication-type="book"><string-name><surname>Poortman</surname>, <given-names>Eva B.</given-names></string-name> <year>2014</year>. <chapter-title>Between intersective and &#8216;split&#8217; interpretations of predicate conjunction: The role of typicality</chapter-title>. In <string-name><given-names>Judith</given-names> <surname>Degen</surname></string-name>, <string-name><given-names>Michael</given-names> <surname>Franke</surname></string-name> &amp; <string-name><given-names>Noah</given-names> <surname>Goodman</surname></string-name> (eds.), <source>Proceedings of the formal &amp; experimental pragmatics workshop</source>, <fpage>36</fpage>&#8211;<lpage>42</lpage>. <publisher-loc>T&#252;bingen</publisher-loc>.</mixed-citation></ref>
<ref id="B13"><label>13</label><mixed-citation publication-type="thesis"><string-name><surname>Poortman</surname>, <given-names>Eva B.</given-names></string-name> <year>2017a</year>. <source>Concepts and plural predication: The effects of conceptual knowledge on the interpretation of reciprocal and conjunctive plural constructions</source>. <publisher-loc>Utrecht</publisher-loc>: <publisher-name>Utrecht University</publisher-name> dissertation.</mixed-citation></ref>
<ref id="B14"><label>14</label><mixed-citation publication-type="book"><string-name><surname>Poortman</surname>, <given-names>Eva B.</given-names></string-name> <year>2017b</year>. <chapter-title>Concept typicality and the interpretation of plural predicate conjunction</chapter-title>. In <string-name><given-names>James A.</given-names> <surname>Hampton</surname></string-name> &amp; <string-name><given-names>Yoad</given-names> <surname>Winter</surname></string-name> (eds.), <source>Compositionality and Concepts in Linguistics and Psychology</source> (Language, Cognition, and Mind) <volume>3</volume>. <fpage>139</fpage>&#8211;<lpage>162</lpage>. <publisher-name>Springer</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1007/978-3-319-45977-6_6</pub-id></mixed-citation></ref>
<ref id="B15"><label>15</label><mixed-citation publication-type="journal"><string-name><surname>Poortman</surname>, <given-names>Eva B.</given-names></string-name> &amp; <string-name><given-names>Liina</given-names> <surname>Pylkk&#228;nen</surname></string-name>. <year>2016</year>. <article-title>Adjective conjunction as a window into the LATL&#8217;s contribution to conceptual combination</article-title>. <source>Brain and Language</source> <volume>160</volume>. <fpage>50</fpage>&#8211;<lpage>60</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.bandl.2016.07.006</pub-id></mixed-citation></ref>
<ref id="B16"><label>16</label><mixed-citation publication-type="journal"><string-name><surname>Quen&#233;</surname>, <given-names>Hugo</given-names></string-name> &amp; <string-name><given-names>Huub</given-names> <surname>van den Bergh</surname></string-name>. <year>2008</year>. <article-title>Examples of mixed-effects modeling with crossed random effects and with binomial data</article-title>. <source>Journal of Memory and Language</source> <volume>59</volume>(<issue>4</issue>). <fpage>413</fpage>&#8211;<lpage>425</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.jml.2008.02.002</pub-id></mixed-citation></ref>
<ref id="B17"><label>17</label><mixed-citation publication-type="thesis"><string-name><surname>Roberts</surname>, <given-names>Craige</given-names></string-name>. <year>1987</year>. <source>Modal subordination, anaphora, and distributivity</source>. <publisher-loc>Amherst, MA</publisher-loc>: <publisher-name>University of Massachusetts Amherst</publisher-name> dissertation.</mixed-citation></ref>
<ref id="B18"><label>18</label><mixed-citation publication-type="book"><string-name><surname>Rosch</surname>, <given-names>Eleanor</given-names></string-name>. <year>1973</year>. <chapter-title>On the internal structure of perceptual and semantic categories</chapter-title>. In <string-name><given-names>Timothy E.</given-names> <surname>Moore</surname></string-name> (ed.), <source>Cognitive development and the acquisition of language</source>, <fpage>111</fpage>&#8211;<lpage>144</lpage>. <publisher-loc>New York</publisher-loc>: <publisher-name>Academic Press</publisher-name>.</mixed-citation></ref>
<ref id="B19"><label>19</label><mixed-citation publication-type="journal"><string-name><surname>Rosch</surname>, <given-names>Eleanor</given-names></string-name> &amp; <string-name><given-names>Carolyn B.</given-names> <surname>Mervis</surname></string-name>. <year>1975</year>. <article-title>Family resemblances: Studies in the internal structure of categories</article-title>. <source>Cognitive Psychology</source> <volume>7</volume>. <fpage>573</fpage>&#8211;<lpage>605</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/0010-0285(75)90024-9</pub-id></mixed-citation></ref>
<ref id="B20"><label>20</label><mixed-citation publication-type="journal"><string-name><surname>Sabato</surname>, <given-names>Sivan</given-names></string-name> &amp; <string-name><given-names>Yoad</given-names> <surname>Winter</surname></string-name>. <year>2012</year>. <article-title>Relational domains and the interpretation of reciprocals</article-title>. <source>Linguistics and Philosophy</source> <volume>25</volume>(<issue>3</issue>). <fpage>191</fpage>&#8211;<lpage>241</lpage>. DOI: <pub-id pub-id-type="doi">10.1007/s10988-012-9117-x</pub-id></mixed-citation></ref>
<ref id="B21"><label>21</label><mixed-citation publication-type="journal"><string-name><surname>Scontras</surname>, <given-names>Gregory</given-names></string-name> &amp; <string-name><given-names>Noah D.</given-names> <surname>Goodman</surname></string-name>. <year>2017</year>. <article-title>Resolving uncertainty in plural predication</article-title>. <source>Cognition</source> <volume>168</volume>. <fpage>294</fpage>&#8211;<lpage>311</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.cognition.2017.07.002</pub-id></mixed-citation></ref>
<ref id="B22"><label>22</label><mixed-citation publication-type="journal"><string-name><surname>Smith</surname>, <given-names>Edward E.</given-names></string-name>, <string-name><given-names>Edward J.</given-names> <surname>Shoben</surname></string-name> &amp; <string-name><given-names>Lance J.</given-names> <surname>Rips</surname></string-name>. <year>1974</year>. <article-title>Structure and process in semantic memory: A featural model for semantic decisions</article-title>. <source>Psychological Review</source> <volume>81</volume>(<issue>3</issue>). <fpage>214</fpage>&#8211;<lpage>241</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/h0036351</pub-id></mixed-citation></ref>
<ref id="B23"><label>23</label><mixed-citation publication-type="book"><string-name><surname>Winter</surname>, <given-names>Yoad</given-names></string-name>. <year>2017</year>. <chapter-title>Critical typicality: truth judgements and compositionality with plurals and other gradable concepts</chapter-title>. In <string-name><given-names>James A.</given-names> <surname>Hampton</surname></string-name> &amp; <string-name><given-names>Yoad</given-names> <surname>Winter</surname></string-name> (eds.), <source>Compositionality and Concepts in Linguistics and Psychology</source> (Language, Cognition, and Mind) <volume>3</volume>. <fpage>163</fpage>&#8211;<lpage>190</lpage>. <publisher-name>Springer</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1007/978-3-319-45977-6_7</pub-id></mixed-citation></ref>
</ref-list>
</back>
</article>