<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.1" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">2397-1835</journal-id>
<journal-title-group>
<journal-title>Glossa: a journal of general linguistics</journal-title>
</journal-title-group>
<issn pub-type="epub">2397-1835</issn>
<publisher>
<publisher-name>Ubiquity Press</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5334/gjgl.708</article-id>
<article-categories>
<subj-group>
<subject>Research</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>The Nordic research infrastructure for syntactic variation: Possibilities, limitations and achievements</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<contrib-id contrib-id-type="orcid">http://orcid.org/0000-0003-1017-8657</contrib-id>
<name>
<surname>Vangsnes</surname>
<given-names>&#216;ystein Alexander</given-names>
</name>
<email>oystein.vangsnes@uit.no</email>
<xref ref-type="aff" rid="aff-1">1</xref>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Johannessen</surname>
<given-names>Janne Bondi</given-names>
</name>
<xref ref-type="aff" rid="aff-3">3</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>CASTL and AcqVA, Department of Language and Culture, UiT The Arctic University of Norway, Langnes, 9037 Troms&#248;, NO</aff>
<aff id="aff-2"><label>2</label>Department of Language, Literature, Mathematics and Interpreting, Western Norway University of Applied Sciences, NO</aff>
<aff id="aff-3"><label>3</label>MultiLing, ILN, University of Oslo, Blindern N-0317, Oslo, NO</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2019-02-08">
<day>08</day>
<month>02</month>
<year>2019</year>
</pub-date>
<pub-date pub-type="collection">
<year>2019</year>
</pub-date>
<volume>4</volume>
<issue>1</issue>
<elocation-id>26</elocation-id>
<history>
<date date-type="received" iso-8601-date="2018-05-31">
<day>31</day>
<month>05</month>
<year>2018</year>
</date>
<date date-type="accepted" iso-8601-date="2018-11-13">
<day>13</day>
<month>11</month>
<year>2018</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2019 The Author(s)</copyright-statement>
<copyright-year>2019</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.glossa-journal.org/articles/10.5334/gjgl.708/"/>
<abstract>
<p>The Scandinavian Dialect Syntax project was a collaboration between ten research groups from all of the five Nordic countries lasting for a period of about ten years. Besides resulting in a large number of scientific papers and theses on a range of different topics, a concrete outcome of the collaboration was the establishment of lasting research infrastructures in terms of two databases: the Nordic Dialect Corpus (NDC) and the Nordic Syntax Database (NSD). This paper first describes the two infrastructures and then proceeds to showcase how they may be used for the exploration of two selected dialect syntactic topics: the relative placement of sentence adverbs and infinitive markers (&#177;split infinitives) across varieties of Mainland North Germanic, and the lack of Verb Second in <italic>wh-</italic>questions across Norwegian dialects.</p>
</abstract>
<kwd-group>
<kwd>dialects</kwd>
<kwd>syntax</kwd>
<kwd>split infinitives</kwd>
<kwd><italic>wh</italic>-questions</kwd>
<kwd>Scandinavian</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec>
<title>1 Introduction</title>
<p>Since 2003 and for a period of about ten years a network of ten research groups in the five Nordic countries worked within the Scandinavian Dialect Syntax project (ScanDiaSyn) towards mapping syntactic variation across the North Germanic dialect continuum. Two major research tools grew out of the collaboration: the Nordic Dialect Corpus (NDC) and the Nordic Syntax Database (NSD), see Section 2.</p>
<p>In 2014 <xref ref-type="bibr" rid="B19"><italic>The Nordic Atlas of Language Structures Online (NALS)</italic></xref> was launched with some 50 papers based on data from the two databases organised under the following six general topics: (i) Noun phrases, (ii) Verb phrase: Argument structure and verb particles, (iii) Verb placement, (iv) Middle field/TP: Subject placement, object shift, auxiliaries and tense marking, (v) Left and right periphery: complementisers, questions, extractions etc., and (vi) Binding and co-reference. NALS is formally organised as a journal, and has continued afterwards to publish papers that map variation across the North Germanic dialect continuum.</p>
<p>In this paper we will discuss two syntactic phenomena on the basis of data available from the dual NDC/NSD research infrastructure: (i) the relative placement of adverbs and infinitive markers (&#177;split infinitive), and (ii) non-V2 in matrix <italic>wh</italic>-questions across Norwegian dialects. For each of these we discuss possibilities, limitations and achievements posed by the research infrastructure as promised by the title of the paper. Section 2 presents the two types of research infrastructure. Section 3 presents the investigation of placement of adverbs and infinitives, and also illustrates how the corpus and database can be used to find the empirical evidence needed. Section 4 contains a thorough presentation of the variation of word order in <italic>wh</italic>-questions, while Section 5 concludes the paper.</p>
</sec>
<sec>
<title>2 The two research infrastructures</title>
<p>The <xref ref-type="bibr" rid="B26">ScanDiaSyn network</xref> of research groups consisted of a number of smaller and bigger projects funded by a variety of sources. In effect, ScanDiaSyn was therefore a project umbrella and there was also funding for the network itself from the Nordic research bodies NordForsk and NOS-HS. In the various countries national funding was obtained to carry out basic and systematic data collection in the projects DanDiaSyn, FinDiaSyn, IceDiaSyn, NorDiaSyn, and SweDiaSyn. The Norwegian national project was responsible for the technical solutions, including building the Nordic Dialect Corpus (<xref ref-type="bibr" rid="B11">Johannessen et al. 2009</xref>, <xref ref-type="bibr" rid="B13">2014</xref>) and the Nordic Syntax Database (<xref ref-type="bibr" rid="B15">Lindstad et al. 2009</xref>) as well as getting the annotations (double transcriptions and tagging) done (<xref ref-type="bibr" rid="B10">Johannessen 2017</xref>). Furthermore, between 2005 and 2010 the ScanDiaSyn umbrella also included substantial funding for the Nordic Center of Excellence in Microcomparative Syntax (NORMS), which financed a number of postdoctoral visiting fellows, cross-institutional thematic research groups, fieldwork trips to targeted areas in the Nordic countries, as well as seminars and conferences. Between 2005 and 2010 a project blog was kept up with entries written both in Scandinavian and English (mainly), including quite extensive reports from the NORMS field trips, and these reports can be found at <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://tinyurl.com/NORMSfieldwork">https://tinyurl.com/NORMSfieldwork</ext-link>.</p>
<p>In total, data from 228 different locations across the Nordic countries were collected in the project, and the data were of two types: (i) recordings of spontaneous speech from conversations between and interviews with dialect speakers, and (ii) judgment data on a list of prepared test sentences that probe a number of different syntactic constructions. These data formed the basis for the two different infrastructures in the project: the <italic><xref ref-type="bibr" rid="B21">Nordic Dialect Corpus</xref></italic> and the <italic><xref ref-type="bibr" rid="B22">Nordic Syntax Database</xref></italic>.</p>
<p>The Nordic Dialect Corpus (NDC) consists of spontaneous speech. Ideally, each location in the corpus should be represented by four speakers (two age groups and both genders), and each speaker would be recorded both in an interview and a free conversation. However, in practice, the different funding situations in the different countries meant that this could not always be achieved. For example, in Sweden, there was less funding, but fortunately the project was given access to interview data from a recently completed project, SweDia 2000. On the other end of the scale was Norway, where funding covered extensive fieldwork across the whole country in addition to the development of the technical research infrastructure. The NDC corpus contains transcribed recordings with over 3 Million words, by 823 dialect speakers from 228 locations in six countries (Denmark, Faroe Islands, Finland, Iceland, Norway and Sweden), mainly sampled between 2005 and 2010.</p>
<p>The Nordic Syntax Database (NSD) contains judgment data for a number of test sentences (140&#8211;240, depending on country) probing variation in various (morpho)syntactic phenomena from more than 900 informants &#8211; many of the same ones as in the NDC. The sentences were carefully selected to represent grammatical phenomena that we knew varied across dialects in one or more of the five languages. A major goal was to present sentences even in locations where a phenomenon was thought not to exist. This way we would obtain negative as well as positive data, making it possible to draw new syntactic isoglosses across the North Germanic dialect continuum.</p>
<p>The locations represented in the infrastructure were chosen to ensure a good geographical distribution and also to cover well-known local/regional dialect boundaries, and although most locations are in rural areas there are also some cities in the Norwegian and Danish list of locations. For the Swedish speaking area, the SweDia 2000 material (see above) only included material from rural and semi-rural locations, and hence no Swedish (and Finnish) cities are represented in the infrastructure. Since geographical distribution was prioritised, the locations were not balanced for population size, hence there are both &#8220;small&#8221; and &#8220;big&#8221; dialects in the sample.</p>
<p>The fieldworkers were a good mix of senior and junior researchers and student assistants who travelled to the locations of the informants to carry out the data collection. In most cases two fieldworkers would do the data sampling together. In some cases the fieldworkers would speak a similar dialect as the informants, but in other cases not.</p>
<p>The test sentences in the questionnaire were pre-recorded by a speaker of the same regional variety so that the pronunciation would be the same as or similar to that of the informants. This was done to avoid the sentences being deemed unacceptable because of pronunciation. The informants were instructed to judge the sentences on a Likert scale from 1 (bad) to 5 (good) according to their own dialect intuitions, and they would give their judgments after hearing each pre-recorded sentence. These questionnaire sessions lasted about one to one and a half hours. The recording sessions consisted of an interview of about 15&#8211;20 minutes with one of the fieldworkers and a conversation with another informant of about 30&#8211;45 minutes. Sampling data from four informants at one location normally took a full work day.</p>
<p>The informants were typically recruited through a local contact person according to a set of criteria targeting traditional dialect speakers. Beyond age and gender information sociolinguistic information about the informants was not recorded. The informants were not paid, but given a symbolic gift as a token of gratitude and they were also served coffee, tea and soft drinks as well as fruit and (non-crunchy) candy at the sessions. For further details about the project logistics, including methodologies, see Vangsnes (<xref ref-type="bibr" rid="B32">2007a</xref>; <xref ref-type="bibr" rid="B33">b</xref>), Johannessen et al. (<xref ref-type="bibr" rid="B12">2008</xref>), Lindstad et al. (<xref ref-type="bibr" rid="B15">2009</xref>), Johannessen et al. (<xref ref-type="bibr" rid="B13">2014</xref>).</p>
<p>The papers in the <xref ref-type="bibr" rid="B20"><italic>NALS Journal</italic></xref>, which typically exploit both the the NDC (corpus) and the NSD (judgment database), show that the dialect infrastructure developed under the ScanDiaSyn umbrella does indeed allow researchers to investigate new isoglosses and dialect phenomena across the Nordic countries. The two case studies to be presented in this paper should serve to make the same point.</p>
</sec>
<sec>
<title>3 Case study 1: Relative placement of infinitival marker and negation (&#177;split infinitives)</title>
<p>The differing relative placement of infinitival markers and adverbs across the Scandinavian written languages is a well-known issue (see for example <xref ref-type="bibr" rid="B8">Hulth&#233;n 1947</xref>; <xref ref-type="bibr" rid="B4">Faarlund et al. 1997</xref>). The received wisdom is that Danish requires the adverb (especially the negative adverb) to be placed before the infinitival marker, as in (1), whereas Swedish requires the adverb to follow the infinitival marker, as in (2), and Norwegian is supposed to accept both orders. The &#8220;Danish pattern&#8221; is what often is referred to as &#8220;unsplit infinitives&#8221; since the adverb does not split the infinitival marker from the verb, and conversely the &#8220;Swedish pattern&#8221; is generally referred to as &#8220;split infinitives&#8221; precisely since the adverb separates the infinitival marker from the verb. A third pattern, given in (3), is the standard word order in Icelandic infinitivals (with an infinitival marker), and this pattern was also tested for the mainland languages (and Faroese). The sentences are all given here with Bokm&#229;l Norwegian words and orthography.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(1)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Kjell</p></list-item>
<list-item><p>kjell</p></list-item>
</list>
<list list-type="word">
<list-item><p>hadde</p></list-item>
<list-item><p>had</p></list-item>
</list>
<list list-type="word">
<list-item><p>lenge</p></list-item>
<list-item><p>long</p></list-item>
</list>
<list list-type="word">
<list-item><p>pr&#248;vd</p></list-item>
<list-item><p>tried</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>ikke</bold></p></list-item>
<list-item><p>not</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>&#229;</bold></p></list-item>
<list-item><p>to</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>komme</bold></p></list-item>
<list-item><p>come</p></list-item>
</list>
<list list-type="word">
<list-item><p>for</p></list-item>
<list-item><p>too</p></list-item>
</list>
<list list-type="word">
<list-item><p>sent</p></list-item>
<list-item><p>late</p></list-item>
</list>
<list list-type="word">
<list-item><p>p&#229;</p></list-item>
<list-item><p>at</p></list-item>
</list>
<list list-type="word">
<list-item><p>jobb.</p></list-item>
<list-item><p>work</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;Kjell had for a long time tried not to come too late for work.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(2)</td>
<td>Kjell hadde lenge pr&#248;vd <bold>&#229; ikke komme</bold> for sent p&#229; jobb.</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(3)</td>
<td>Kjell hadde lenge pr&#248;vd <bold>&#229; komme ikke</bold> for sent p&#229; jobb.</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The issue of &#8220;split infinitive&#8221; is well-known also from English grammar and is a source of controversy where prescriptivists favour the use of unsplit over split infinitives (see <xref ref-type="bibr" rid="B7">Huddleston &amp; Pullum 2002: 581f, for references</xref>). This same prescriptivism can be witnessed also in the context of Norwegian as unsplit infinitives have traditionally been the recommended word order. Furthermore, the Norwegian reference grammar (<xref ref-type="bibr" rid="B4">Faarlund et al. 1997: 997</xref>) claims that while both orders are possible in Norwegian, the unsplit pattern is the most natural for many language users (<xref ref-type="bibr" rid="B4">Faarlund 1997: 997</xref>). We will see below that the NSD database does not support this claim.</p>
<p>Figure <xref ref-type="fig" rid="F1">1</xref> is a screenshot from the NSD of sentence (1) as it was presented to the informants, and we see that the Swedish test sentence is worded differently but nevertheless probes the same word order as its Norwegian and Danish counterparts. Similar adjustments across the language-specific questionnaires were made in several cases, but each &#8220;bundle of test sentences&#8221; probing a specific phenomenon is always given a unique identity in the database, in this case the number 143.</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>The test sentence for unsplit infinitives in Danish, Norwegian and Swedish, assigned number 143 in the Nordic Syntax Database.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65086/"/>
</fig>
<p>Searching the database gives a long list with all the answers of all the informants, divided over several results pages. Each informant would grade this sentence (and all the others presented to them) on a scale from 1 to 5. These results are rendered as colour codes in the database, to enable the researcher to assess the results at a glance, see Figure <xref ref-type="fig" rid="F2">2</xref>.</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Some of the results of the evaluations of sentence no. 143 (the unsplit, &#8220;Danish pattern&#8221;).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65087/"/>
</fig>
<p>Even more visually illustrative are the maps that can be generated from the result page. The results from the sentence evaluations performed across the three countries Denmark, Norway and Sweden show very clear and different results, see Maps <xref ref-type="fig" rid="M1">1</xref>, <xref ref-type="fig" rid="M2">2</xref> and <xref ref-type="fig" rid="M3">3</xref>. A white marker means that a sentence has a mean score of 4 or higher at that geographical location, whereas a black marker means that it has a mean score of 2 or lower. In other words, a white marker indicates acceptability, a black one non-acceptability.</p>
<fig id="M1">
<label>Map 1</label>
<caption>
<p>Unsplit infinitive: neg &#8211; C &#8211; Vinf.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65088/"/>
</fig>
<fig id="M2">
<label>Map 2</label>
<caption>
<p>Split infinitive: C &#8211; neg &#8211; Vinf.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65089/"/>
</fig>
<fig id="M3">
<label>Map 3</label>
<caption>
<p>The &#8220;Icelandic pattern&#8221;: C &#8211; Vinf &#8211; neg.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65090/"/>
</fig>
<p>There is a clear acceptance of the unsplit, &#8220;Danish pattern&#8221; in Denmark. Furthermore, in Sweden the unsplit pattern is quite clearly not accepted, but more strikingly, at the great majority of Norwegian locations the pattern is also dismissed. The split, &#8220;Swedish pattern&#8221;, on the other hand, is accepted not just in Sweden, but also in most of Norway, although there is an enclave in the central part of the country (the Tr&#248;ndelag area) where the split pattern seems to be rejected at several locations. In turn, as Map <xref ref-type="fig" rid="M3">3</xref> makes evident, the &#8220;Icelandic pattern&#8221; with the verb preceding negation is not accepted by anyone in the mainland countries (and not in the Faroe Islands either).</p>
<p>A striking result from the database data, evident from the maps, is that the &#8220;Danish pattern&#8221; is hardly accepted at all by the informants from the Norwegian measure points. Only at ten locations in Norway have the informants accepted it, and these are spread quite evenly across the country.</p>
<p>The general picture thus is very clear. Norwegian dialects generally follow the &#8220;Swedish pattern&#8221;, with negation between the infinitival marker and the infinitive. The &#8220;Danish pattern&#8221; is rejected in most cases, and there is only one place in which the &#8220;Danish pattern&#8221; gets a higher score than the Swedish one. The claim by Faarlund et al. (<xref ref-type="bibr" rid="B4">1997: 997</xref>) that the &#8220;Danish pattern&#8221; is more natural for many Norwegian language users is therefore severely weakened by the data in the NSD.</p>
<p>A note on the data from Denmark is in order. First, at one location in the North of Jutland (&#8220;Vendsyssel&#8221;) both the split and the unsplit patterns are accepted. Pedersen (<xref ref-type="bibr" rid="B23">2017: 44ff</xref>) has looked more closely at these data, and she finds that four informants at this location accepts the test sentence. Furthermore, she also points out that there is at least one informant at all of the other locations in Jutland that accepts the sentence, and at the measure point Eastern Jutland (&#8220;&#216;stjylland&#8221;) three informants do so. Second, Pedersen (op. cit.) shows that the existence of the &#8220;Swedish pattern&#8221; has been mentioned and documented in the dialectological literature also for the insular parts of Denmark. In other words, even in Danish dialects, the &#8220;Danish pattern&#8221; does not seem to be as obligatory as one might think.</p>
<p>On the basis of these considerations, it is worth investigating to what extent data from the corpus of spontaneous speech corroborate the results from the database. Defining a search for the &#8220;Danish pattern&#8221; is, however, not trivial since a negation preceding the infinitival marker may belong to the matrix predicate rather than to the infinitival clause: the example in (4) is ambiguous between a high and low attachment for the negation. This is not the case with the &#8220;Swedish pattern&#8221;, see (5), repeated from (2).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(4)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Kjell</p></list-item>
<list-item><p>Kjell</p></list-item>
</list>
<list list-type="word">
<list-item><p>pr&#248;ver</p></list-item>
<list-item><p>tries</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>ikke</bold></p></list-item>
<list-item><p>not</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>&#229;</bold></p></list-item>
<list-item><p>to</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>komme</bold></p></list-item>
<list-item><p>come</p></list-item>
</list>
<list list-type="word">
<list-item><p>for</p></list-item>
<list-item><p>too</p></list-item>
</list>
<list list-type="word">
<list-item><p>sent</p></list-item>
<list-item><p>late</p></list-item>
</list>
<list list-type="word">
<list-item><p>p&#229;</p></list-item>
<list-item><p>at</p></list-item>
</list>
<list list-type="word">
<list-item><p>jobb.</p></list-item>
<list-item><p>work</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>i. &#8216;Kjell does not try to come too late for work.&#8217;</p></list-item>
<list-item><p>ii. &#8216;Kjell tries not to come too late for work.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(5)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Kjell</p></list-item>
<list-item><p>Kjell</p></list-item>
</list>
<list list-type="word">
<list-item><p>hadde</p></list-item>
<list-item><p>had</p></list-item>
</list>
<list list-type="word">
<list-item><p>lenge</p></list-item>
<list-item><p>long</p></list-item>
</list>
<list list-type="word">
<list-item><p>pr&#248;vd</p></list-item>
<list-item><p>tried</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>&#229;</bold></p></list-item>
<list-item><p>to</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>ikke</bold></p></list-item>
<list-item><p>not</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>komme</bold></p></list-item>
<list-item><p>come</p></list-item>
</list>
<list list-type="word">
<list-item><p>for</p></list-item>
<list-item><p>too</p></list-item>
</list>
<list list-type="word">
<list-item><p>sent</p></list-item>
<list-item><p>late</p></list-item>
</list>
<list list-type="word">
<list-item><p>p&#229;</p></list-item>
<list-item><p>at</p></list-item>
</list>
<list list-type="word">
<list-item><p>jobb.</p></list-item>
<list-item><p>work</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;Kjell had for a long time tried not to come too late for work.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>A search for the string [negation]+[infinitival marker], tailored to the &#8220;Danish pattern&#8221; is formulated as in Figure <xref ref-type="fig" rid="F3">3</xref>. The search specifies that the first word should not be a verb in the past or present tense (to try to avoid the ambiguous pattern exemplified in (4)), followed by the negation word and the infinitival marker.</p>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>Search string for unsplit infinitives (the &#8220;Danish pattern&#8221;) in the Nordic Syntax Database.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65091/"/>
</fig>
<p>This search, when limited to just the Norwegian part of the corpus, gives 13 relevant results. All of them turn out to involve the idiomatic phrase <italic>for ikke &#229;</italic> &#8216;in order not to/to not even&#8217;, i.e. only with this preposition and only in this meaning. An example is given in (6).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(6)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>for</p></list-item>
<list-item><p>for</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>itt</bold></p></list-item>
<list-item><p>not</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>&#229;</bold></p></list-item>
<list-item><p>to</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>snakk</bold></p></list-item>
<list-item><p>speak</p></list-item>
</list>
<list list-type="word">
<list-item><p>om</p></list-item>
<list-item><p>about</p></list-item>
</list>
<list list-type="word">
<list-item><p>syklinga</p></list-item>
<list-item><p>cycling.<sc>DEF</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>her</p></list-item>
<list-item><p>her</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;to not even talk about the cycling around here&#8217; (bjugn_19)</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>Other than these 13, there are zero hits for unsplit infinitives (the &#8220;Danish pattern&#8221;) in Norwegian dialects.</p>
<p>The search for the split &#8220;Swedish pattern&#8221;, on the other hand, gives 60 hits from 41 different locations across all of Norway. All of the hits are relevant, and an example is given in (7).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(7)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Han</p></list-item>
<list-item><p>he</p></list-item>
</list>
<list list-type="word">
<list-item><p>sku</p></list-item>
<list-item><p>should</p></list-item>
</list>
<list list-type="word">
<list-item><p>l&#230;r</p></list-item>
<list-item><p>teach</p></list-item>
</list>
<list list-type="word">
<list-item><p>oss</p></list-item>
<list-item><p>us</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>&#229;</bold></p></list-item>
<list-item><p>to</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>ikkje</bold></p></list-item>
<list-item><p>not</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>ver</bold></p></list-item>
<list-item><p>be</p></list-item>
</list>
<list list-type="word">
<list-item><p>redd</p></list-item>
<list-item><p>afraid</p></list-item>
</list>
<list list-type="word">
<list-item><p>uveret.</p></list-item>
<list-item><p>storm.<sc>DEF</sc></p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;He was going to teach us not to be afraid of the storm.&#8217; (stamsund_03gm)</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The NDC has a map function which allows the user to generate a map to show which locations the hits are from, and Map <xref ref-type="fig" rid="M4">4</xref> shows the distribution of the 60 split infinitives found in Norwegian dialects.</p>
<fig id="M4">
<label>Map 4</label>
<caption>
<p>The 60 hits of split infinitives found in the NDC spread across 41 locations in all of Norway.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65092/"/>
</fig>
<p>There are some important differences between the maps generated by the NSD, Maps <xref ref-type="fig" rid="M1">1</xref>&#8211;<xref ref-type="fig" rid="M3">3</xref>, and the NDC, Map <xref ref-type="fig" rid="M4">4</xref>. The former only generates maps on the basis of sentence evaluations, while the latter generates hits from spontaneously produced speech. This means that while the locations in Map <xref ref-type="fig" rid="M4">4</xref> show where the split infinitive (the &#8220;Swedish pattern&#8221;) has been attested, we cannot draw the conclusion that the missing points on the map are places where it could not occur. We already know that the unsplit infinitive (the &#8220;Danish pattern&#8221;) has only been attested for a sub-construction of all logically possible syntactic possibilities, and therefore that it is not one that would cover the unmarked places on the map. When a corpus does not attest a certain usage, it may be because the informants simply did not use that construction during the recorded conversation session.</p>
<p>The maps illustrate clearly why language research benefits from both a database and a corpus kind of infrastructure. The database data from the NSD has many more hits than the corpus data from NDC. However, a database based on evaluations of sentences can only answer questions that have been asked, and databases will therefore contain only a subset of variations of a construction. A corpus, on the other hand, where informants speak freely, will exemplify many different constructions, even ones that the researcher has not thought of beforehand. At the same time, it is to some extent arbitrary what constructions conversation partners use in recordings. This is exemplified in Map <xref ref-type="fig" rid="M4">4</xref>, which has far fewer locations than Maps <xref ref-type="fig" rid="M1">1</xref>, <xref ref-type="fig" rid="M2">2</xref>, <xref ref-type="fig" rid="M3">3</xref>.</p>
<p>What this investigation shows, is that when the researcher is fortunate enough to have a database of evaluations for a particular structure, there will be hits for all the locations investigated. A corpus does not necessarily cover all locations if a construction is not among the most common ones. Still, the corpus can be used to check whether informants in the database have given answers that are indeed compatible with their own language production. Our investigation in this particular case shows that this is indeed the case. The placement of the adverb with respect to the infinitival marker in Norwegian turns out to follow the Swedish pattern, illustrated in Maps <xref ref-type="fig" rid="M1">1</xref>, <xref ref-type="fig" rid="M2">2</xref>, <xref ref-type="fig" rid="M3">3</xref>, and the production data from the corpus show the same, even more convincingly, in Map <xref ref-type="fig" rid="M4">4</xref>. And the whole infrastructure together shows that the claim made in the Norwegian reference grammar (<xref ref-type="bibr" rid="B4">Faarlund et al. 1997</xref>) is not supported by our empirical investigations.</p>
<p>We now turn to look at a different and far more complex issue, namely the lack of Verb Second in matrix <italic>wh-</italic>questions in Norwegian dialects.</p>
</sec>
<sec>
<title>4 Case study 2: Non-V2 in matrix <italic>wh</italic>-questions across Norwegian dialects</title>
<sec>
<title>4.1 Previous knowledge</title>
<p>The traditional portrayal of Norwegian and the North Germanic languages in general is that they are well-behaved Verb Second languages, i.e. with the finite verb in a fronted Wackernagel position in matrix clauses, always occurring before the subject whenever a non-subject introduces the clause. This is exemplified by the declarative clause in (8), and the comparison with the idiomatic English translation serves to make the point.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(8)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>&#160;&#160;I morgon</p></list-item>
<list-item><p>&#160;&#160;tomorrow</p></list-item>
</list>
<list list-type="word">
<list-item><p>skal</p></list-item>
<list-item><p>will</p></list-item>
</list>
<list list-type="word">
<list-item><p>studentane</p></list-item>
<list-item><p>students.<sc>DEF</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>ta</p></list-item>
<list-item><p>take</p></list-item>
</list>
<list list-type="word">
<list-item><p>eksamen.</p></list-item>
<list-item><p>exam</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#160;&#160;&#8216;Tomorrow the students will take the exam.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>*I morgon studentane skal ta eksamen</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>In Lohndal et al. (<xref ref-type="bibr" rid="B16">forthcoming</xref>) an overview of exceptions to the V2 requirement in Norwegian is given. The exceptions are more than one tends to acknowledge in general descriptions of the language. In short, the main message in the paper is that Verb Second in Norwegian cannot be an effect of a single macro-parameter, but is rather due to several minor rules which conspire to give the impression that V2 is almost omnipresent (cf. also Weerman 1989).</p>
<p>One phenomenon which all the same has received considerable attention is the lack of V2 in matrix <italic>wh</italic>-questions, see e.g. Iversen (<xref ref-type="bibr" rid="B9">1918</xref>), Elstad (<xref ref-type="bibr" rid="B3">1982</xref>), Nordg&#229;rd (<xref ref-type="bibr" rid="B18">1985</xref>), &#197;farli (<xref ref-type="bibr" rid="B1">1986</xref>), Taraldsen (<xref ref-type="bibr" rid="B29">1986</xref>), Lie (<xref ref-type="bibr" rid="B14">1992</xref>), Fiva (<xref ref-type="bibr" rid="B5">1996</xref>), Nilsen (<xref ref-type="bibr" rid="B17">1996</xref>), Westergaard (<xref ref-type="bibr" rid="B37">2003</xref>; <xref ref-type="bibr" rid="B38">2005</xref>; <xref ref-type="bibr" rid="B39">2009a</xref>; <xref ref-type="bibr" rid="B40">b</xref>; <xref ref-type="bibr" rid="B41">2017</xref>), Westergaard &amp; Vangsnes (<xref ref-type="bibr" rid="B42">2005</xref>), Vangsnes (<xref ref-type="bibr" rid="B31">2005</xref>), Rognes (<xref ref-type="bibr" rid="B25">2011</xref>), Reite (<xref ref-type="bibr" rid="B24">2011</xref>), Vangsnes &amp; Westergaard (<xref ref-type="bibr" rid="B34">2014</xref>), Westendorp (<xref ref-type="bibr" rid="B35">2017</xref>; <xref ref-type="bibr" rid="B36">2018</xref>), and Westergaard, Vangsnes &amp; Lohndal (<xref ref-type="bibr" rid="B43">2017</xref>). Iversen (<xref ref-type="bibr" rid="B9">1918: 37</xref>) is an early source commenting on the phenomenon. In his study of the syntax of the city dialect of Troms&#248;, he notes that the interrogative pronouns <italic>k&#230;m</italic> &#8216;who&#8217; and <italic>ka</italic> &#8216;what&#8217; are associated with &#8220;a quaint word order&#8221; in that direct questions (i.e. matrix ones) show the same word order as indirect questions (i.e. embedded ones) with these <italic>wh</italic>-pronouns.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(9)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><italic>Troms&#248; dialect</italic> (<xref ref-type="bibr" rid="B9">Iversen 1918</xref>)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="word">
<list-item><p>K&#230;m</p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>tr&#230;fte?</p></list-item>
<list-item><p>met</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;Who did you meet?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Ka</p></list-item>
<list-item><p>what</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>hete?</p></list-item>
<list-item><p>are-called</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;What is your name?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The &#8216;quaintness&#8217; of these structures is, of course, that the standard written language would require V2 in the corresponding cases, exemplified here by standard Nynorsk Norwegian examples.</p>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(10)</td>
<td>a.</td>
<td>Kven trefte du?/*Kven du trefte?</td>
</tr>
<tr>
<td>&#160;</td>
<td>b.</td>
<td>Kva heiter du?/*Kva du heiter?</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>At the outset of the <xref ref-type="bibr" rid="B26">ScanDiaSyn network</xref> it was already well established that this lack of Verb Second is widespread across Norwegian dialects. The only area for which no instance of the phenomenon had been reported seemed to be Central Eastern Norway around the capital Oslo and the adjacent coastal areas to the south. It had also been established that there is considerable variation across the dialects. The first four points below are basic and shared characteristics of the phenomenon.</p>
<list list-type="alpha-lower">
<list-item><p>The finite verb is in a position to the right of sentence adverbs (i.e. has not moved to C).</p></list-item>
<list-item><p>In subject <italic>wh</italic>-questions the complementiser <italic>som</italic> appears in second position (and the finite verb in a position to the right of sentence adverbs).</p></list-item>
<list-item><p>All dialects that allow non-V2 in <italic>wh</italic>-questions, also allow V2: the choice of &#177;V2 appears to be governed by information structure in that V2 is preferred when the subject is given information and non-V2 when the subject is new information.</p></list-item>
<list-item><p>Dialects that allow non-V2 in matrix <italic>wh</italic>-questions also allow the complementiser <italic>som</italic> before the trace position of an extracted <italic>wh</italic>-subject, hence seemingly violating a COMP trace effect.</p></list-item>
</list>
<p>Points a. and b. entail that there is a strong parallelism between matrix <italic>wh-</italic>questions with non-V2 and embedded <italic>wh</italic>-questions: in embedded <italic>wh</italic>-questions the finite verb also appears to the right of a sentence adverb and the appearance of the complementiser <italic>som</italic> is obligatory in subject questions.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(11)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Eg</p></list-item>
<list-item><p>I</p></list-item>
</list>
<list list-type="word">
<list-item><p>lurer</p></list-item>
<list-item><p>wonder</p></list-item>
</list>
<list list-type="word">
<list-item><p>p&#229; &#8230;</p></list-item>
<list-item><p>on</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>&#8230;</p></list-item>
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="word">
<list-item><p>kven</p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>eigentleg</p></list-item>
<list-item><p>actually</p></list-item>
</list>
<list list-type="word">
<list-item><p>trefte./&#8230;</p></list-item>
<list-item><p>met</p></list-item>
</list>
<list list-type="word">
<list-item><p>*kven</p></list-item>
<list-item><p>&#160;&#160;who</p></list-item>
</list>
<list list-type="word">
<list-item><p>trefte</p></list-item>
<list-item><p>met</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>eigentleg.</p></list-item>
<list-item><p>actually</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;I wonder who you actually met.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>&#8230;</p></list-item>
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="word">
<list-item><p>kven</p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p>som</p></list-item>
<list-item><p><sc>SOM</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>eigentleg</p></list-item>
<list-item><p>actually</p></list-item>
</list>
<list list-type="word">
<list-item><p>bestemmer./&#8230;</p></list-item>
<list-item><p>decides</p></list-item>
</list>
<list list-type="word">
<list-item><p>*kven</p></list-item>
<list-item><p>&#160;&#160;who</p></list-item>
</list>
<list list-type="word">
<list-item><p>bestemmer</p></list-item>
<list-item><p>decides</p></list-item>
</list>
<list list-type="word">
<list-item><p>eigentleg.</p></list-item>
<list-item><p>actually</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;I wonder who actually decides.&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The complementiser <italic>som</italic> has no one-to-one English equivalent: it also appears in relative clauses and clefts (corresponding to <italic>that</italic>) and in small clauses and comparatives (corresponding to <italic>as</italic>) (see <xref ref-type="bibr" rid="B28">Stroh-Wollin 2002</xref>; <xref ref-type="bibr" rid="B30">Vangsnes 2004: 22f</xref> for details).</p>
<p>The insight in point c. is due to Westergaard (<xref ref-type="bibr" rid="B37">2003</xref>; <xref ref-type="bibr" rid="B38">2005</xref>), who studied V2 vs. non-V2 quantitatively in a corpus of the Troms&#248; dialect. Point d. can be credited to Nordg&#229;rd (<xref ref-type="bibr" rid="B18">1985</xref>) and is illustrated by the example in (12).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(12)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Kem</p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>trur</p></list-item>
<list-item><p>think</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>som</bold></p></list-item>
<list-item><p><sc>SOM</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>___</bold></p></list-item>
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="word">
<list-item><p>e</p></list-item>
<list-item><p>is</p></list-item>
</list>
<list list-type="word">
<list-item><p>i</p></list-item>
<list-item><p>in</p></list-item>
</list>
<list list-type="word">
<list-item><p>baren?</p></list-item>
<list-item><p>bar.<sc>DEF</sc></p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;Who do you think is in the bar?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The following four points concern issues of <italic>variation</italic> across the dialects that allow non-V2 in matrix <italic>wh</italic>-questions in the first place.</p>
<list list-type="simple">
<list-item><p>e. Many dialects allow non-V2 only with the short, monosyllabic items <italic>kven</italic> &#8216;who&#8217;, <italic>kva</italic> &#8216;what&#8217; and <italic>kor</italic> &#8216;where&#8217;.</p></list-item>
<list-item><p>f. Some dialects also allow complex <italic>wh-</italic>items and <italic>wh</italic>-phrases with non-V2.</p></list-item>
<list-item><p>g. Some dialects allow non-V2 only in subject <italic>wh</italic>-questions. i.e. with <italic>som</italic> in second position.</p></list-item>
<list-item><p>h. Some dialects allow complex <italic>wh</italic>-subjects but only simple non-subject <italic>wh</italic>-constituents, hence make a &#177;subject distinction with regard to the complexity of the <italic>wh</italic>-constituent.</p></list-item>
</list>
<p>Point e. was noted for the Troms&#248; city dialect already by Iversen (<xref ref-type="bibr" rid="B9">1918: 37</xref>), and stated more broadly as a trait of Northern Norwegian dialects by Elstad (<xref ref-type="bibr" rid="B3">1982</xref>). Point f. was shown by Nordg&#229;rd (<xref ref-type="bibr" rid="B18">1985</xref>) and &#197;farli (<xref ref-type="bibr" rid="B1">1986</xref>) for Northwestern Norwegian dialects. Point g. is due to Lie (<xref ref-type="bibr" rid="B14">1992: 66</xref>) who noted that some Western Norwegian dialects seem to only allow non-V2 with <italic>wh-</italic>subjects and insertion of <italic>som</italic>. Point h. can be attributed to Fiva (1995) who reported that in a survey of the Troms&#248; dialect many informants found complex <italic>wh</italic>-subjects acceptable followed by <italic>som</italic> but would still only accept short <italic>wh</italic>-constituents in non-subject questions with non-V2.</p>
</sec>
<sec>
<title>4.2 The questionnaire data</title>
<p>These various bits of knowledge informed the development of test sentences for the questionnaire to be used in the project. Since the questionnaire was to probe a long list of different constructions, some corners inevitably had to be cut. One of them was to check if the informants allowed both V2 and non-V2 in matrix <italic>wh</italic>-questions, and along with that, to what extent the choice was governed by information structure. In the Norwegian version of the questionnaire, which at the outset had about 130 sentences, we ended up with the following four sentences regarding non-V2 in matrix <italic>wh</italic>-questions.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(13)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p><bold>Ka</bold></p></list-item>
<list-item><p>what</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>hete?</p></list-item>
<list-item><p>are.called</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;What are you called?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(14)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p><bold>Ka</bold></p></list-item>
<list-item><p>what</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>tid</bold></p></list-item>
<list-item><p>time</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>gikk</p></list-item>
<list-item><p>went</p></list-item>
</list>
<list list-type="word">
<list-item><p>ut</p></list-item>
<list-item><p>out</p></list-item>
</list>
<list list-type="word">
<list-item><p>av</p></list-item>
<list-item><p>of</p></list-item>
</list>
<list list-type="word">
<list-item><p>ungdomsskolen?</p></list-item>
<list-item><p>secondary.school.<sc>DEF</sc></p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;When did you graduate from secondary school?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(15)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p><bold>Kem</bold></p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>som</bold></p></list-item>
<list-item><p><sc>SOM</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>s&#230;ll</p></list-item>
<list-item><p>sells</p></list-item>
</list>
<list list-type="word">
<list-item><p>fiskeutstyr</p></list-item>
<list-item><p>fishing.gear</p></list-item>
</list>
<list list-type="word">
<list-item><p>her</p></list-item>
<list-item><p>here</p></list-item>
</list>
<list list-type="word">
<list-item><p>i</p></list-item>
<list-item><p>in</p></list-item>
</list>
<list list-type="word">
<list-item><p>bygda?</p></list-item>
<list-item><p>village.<sc>DEF</sc></p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;Who&#8217;s selling fishing gear in this village?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(16)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p><bold>Kor</bold></p></list-item>
<list-item><p>how</p></list-item>
</list>
<list list-type="word">
<list-item><p><bold>mange</bold></p></list-item>
<list-item><p>many</p></list-item>
</list>
<list list-type="word">
<list-item><p>eleva</p></list-item>
<list-item><p>pupils</p></list-item>
</list>
<list list-type="word">
<list-item><p>som</p></list-item>
<list-item><p><sc>SOM</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>g&#229;r</p></list-item>
<list-item><p>go</p></list-item>
</list>
<list list-type="word">
<list-item><p>p&#229;</p></list-item>
<list-item><p>on</p></list-item>
</list>
<list list-type="word">
<list-item><p>den</p></list-item>
<list-item><p>the</p></list-item>
</list>
<list list-type="word">
<list-item><p>her</p></list-item>
<list-item><p>here</p></list-item>
</list>
<list list-type="word">
<list-item><p>skolen?</p></list-item>
<list-item><p>school.<sc>DEF</sc></p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;How many pupils go to this school?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>(13) has a short <italic>wh-</italic>predicative, (14) has a complex <italic>wh</italic>-adverb, (15) has a short <italic>wh</italic>-subject and (16) has a complex <italic>wh</italic>-subject, and in sum these sentences should serve to probe the issue of short versus complex <italic>wh</italic>-constituents, the &#177;subject condition and, furthermore, whether subject and non-subject questions differ with respect to allowing complex <italic>wh</italic>-constituents with non-V2.</p>
<p>However, the sentences would not serve to detect whether there would be a difference between arguments and non-arguments, or if different <italic>wh</italic>-adverbs would give different results. After the data collection had begun, an additional test sentence with a different <italic>wh-</italic>adverb was added to the questionnaire in order to possibly obtain more information about other relevant factors.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(17)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p><bold>Koffer</bold></p></list-item>
<list-item><p>what-for</p></list-item>
</list>
<list list-type="word">
<list-item><p>han</p></list-item>
<list-item><p>he</p></list-item>
</list>
<list list-type="word">
<list-item><p>va</p></list-item>
<list-item><p>was</p></list-item>
</list>
<list list-type="word">
<list-item><p>s&#229;</p></list-item>
<list-item><p>so</p></list-item>
</list>
<list list-type="word">
<list-item><p>sur,</p></list-item>
<list-item><p>grumpy</p></list-item>
</list>
<list list-type="word">
<list-item><p>egentli?</p></list-item>
<list-item><p>actually</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>&#8216;Why was he so grumpy?&#8217; (&#8216;What was the actual reason for his grumpiness?&#8217;)</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The database results for the four initial non-V2 <italic>wh</italic>-questions have recently been presented in Westergaard et al. (2012; <xref ref-type="bibr" rid="B43">2017</xref>). The results by and large confirm the findings reported in the earlier literature, including statements e. to h. above, but now on the basis of a much more comprehensive and systematic data collection (four informants from 107 locations spread out across Norway). Four types of dialects allowing non-V2 emerge from the NSD data:</p>
<list list-type="alpha-upper">
<list-item><p>Complex <italic>wh</italic>-constituents are allowed with non-V2 in both subject and non-subject questions (i.e. all four test sentences)</p></list-item>
<list-item><p>Only short <italic>wh</italic>-constituents are allowed with non-V2 in subject and non-subject questions alike (i.e. (13) and (15)).</p></list-item>
<list-item><p>Complex <italic>wh</italic>-subjects are allowed with non-V2, but only short <italic>wh</italic>-constituents are allowed in non-subject questions (i.e. all but (14) are accepted).</p></list-item>
<list-item><p>Only <italic>wh</italic>-subjects are allowed with non-V2 (i.e. (15) and (16) are accepted).</p></list-item>
</list>
<p>The distribution of these dialect types is given in Map <xref ref-type="fig" rid="M5">5</xref> which is also published in Westergaard et al. (<xref ref-type="bibr" rid="B43">2017: 26</xref>).</p>
<fig id="M5">
<label>Map 5</label>
<caption>
<p>Map taken from Westergaard et al. (<xref ref-type="bibr" rid="B43">2017</xref>). The distribution of different types of grammars for matrix <italic>wh</italic>-questions across Norway &#8211; A = allows both simple and complex <italic>wh</italic>-constituents with non-V2, B = allows only simple <italic>wh</italic>-constituents with non-V2, C = allows both simple and complex <italic>wh</italic>-subjects but only simple non-subjects with non-V2, D = allows only subject <italic>wh</italic>-questions with non-V2, ? = unclear pattern(s), * = no non-V2 is allowed.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65093/"/>
</fig>
<p>In the area labeled with &#8216;*&#8217; non-V2 is dismissed by the informants consulted, and for the area labeled with &#8216;?&#8217; no clear pattern emerges as far as the authors can see (see also <xref ref-type="bibr" rid="B35">Westendorp 2017</xref>; <xref ref-type="bibr" rid="B36">2018</xref>, for an assessment of the data from this area).</p>
</sec>
<sec>
<title>4.3 The spontaneous speech data</title>
<p>Vangsnes &amp; Westergaard (<xref ref-type="bibr" rid="B34">2014</xref>) present data regarding the phenomenon based on searches in the Nordic Dialect Corpus (NDC). The searches were optimised to target matrix <italic>wh</italic>-questions, and a gross number of 2273 hits were trimmed down to 1332 relevant ones after fragments, exclamatives and embedded clauses had been sorted out.</p>
<p>The distribution of the remaining relevant examples over different <italic>wh-</italic>items and phrases were as given in Table <xref ref-type="table" rid="T1">1</xref> (<xref ref-type="bibr" rid="B34">Vangsnes &amp; Westergaard 2014: 142</xref>). The figures show three things in particular. First, for the short <italic>wh-</italic>items &#8216;what&#8217;, &#8216;who&#8217; and &#8216;where&#8217;, there is a more or less even distribution between V2 and non-V2. Second, the <italic>when</italic> questions come in a middle position with about every fourth instance having non-V2. Third, for the other adverbial <italic>wh-</italic>items and the <italic>wh-</italic>phrases, very few appear with non-V2.</p>
<table-wrap id="T1">
<label>Table 1</label>
<caption>
<p>V2 and non-V2 in matrix <italic>wh</italic>-questions in the Norwegian part of the Nordic Dialect Corpus by type of <italic>wh</italic>-constituent.</p>
</caption>
<table>
<tr>
<th valign="middle" align="left" style="background-color:#f3f3f4;"></th>
<th valign="middle" align="center" colspan="2" style="background-color:#f3f3f4;">V2</th>
<th valign="middle" align="center" colspan="2" style="background-color:#f3f3f4;">non-V2</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">Total</th>
</tr>
<tr>
<th colspan="6"><hr/></th>
</tr>
<tr>
<th valign="middle" align="left" style="background-color:#f3f3f4;"></th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">n</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">%</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">n</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">%</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">n</th>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td valign="top" align="left"><italic>hva</italic> &#8216;what&#8217;</td>
<td valign="top" align="right">284</td>
<td valign="top" align="right">43%</td>
<td valign="top" align="right">376</td>
<td valign="top" align="right">57%</td>
<td valign="top" align="right">660</td>
</tr>
<tr>
<td valign="top" align="left"><italic>hvem</italic> &#8216;who&#8217;</td>
<td valign="top" align="right">50</td>
<td valign="top" align="right">45%</td>
<td valign="top" align="right">61</td>
<td valign="top" align="right">55%</td>
<td valign="top" align="right">111</td>
</tr>
<tr>
<td valign="top" align="left"><italic>hvor</italic> &#8216;where&#8217;</td>
<td valign="top" align="right">68</td>
<td valign="top" align="right">52.3%</td>
<td valign="top" align="right">62</td>
<td valign="top" align="right">47.7%</td>
<td valign="top" align="right">130</td>
</tr>
<tr>
<td valign="top" align="left"><italic>n&#229;r + hva tid</italic> &#8216;when&#8217;</td>
<td valign="top" align="right">58</td>
<td valign="top" align="right">73.4%</td>
<td valign="top" align="right">21<xref ref-type="fn" rid="n1">1</xref></td>
<td valign="top" align="right">26.6%</td>
<td valign="top" align="right">79</td>
</tr>
<tr>
<td valign="top" align="left"><italic>hvorfor</italic> &#8216;why&#8217;</td>
<td valign="top" align="right">46</td>
<td valign="top" align="right">97.9%</td>
<td valign="top" align="right">1</td>
<td valign="top" align="right">2.1%</td>
<td valign="top" align="right">47</td>
</tr>
<tr>
<td valign="top" align="left"><italic>hvordan</italic> &#8216;how&#8217; (manner)</td>
<td valign="top" align="right">119</td>
<td valign="top" align="right">93.0%</td>
<td valign="top" align="right">9</td>
<td valign="top" align="right">7.0%</td>
<td valign="top" align="right">128</td>
</tr>
<tr>
<td valign="top" align="left">&#8216;<italic>wh-</italic>XP&#8217;</td>
<td valign="top" align="right">169</td>
<td valign="top" align="right">95.3%</td>
<td valign="top" align="right">8</td>
<td valign="top" align="right">4.7%</td>
<td valign="top" align="right">177</td>
</tr>
<tr>
<td valign="top" align="left">Total</td>
<td valign="top" align="right">794</td>
<td valign="top" align="right">59.6%</td>
<td valign="top" align="right">538</td>
<td valign="top" align="right">40.4%</td>
<td valign="top" align="right">1332</td>
</tr>
</table>
</table-wrap>
<p>Concerning the first observation, the distribution of V2 versus non-V2 varies across different parts of the country. Vangsnes &amp; Westergaard (<xref ref-type="bibr" rid="B34">2014: 143</xref>) show that for the three short <italic>wh</italic>-items &#8216;what&#8217;, &#8216;who&#8217;, and &#8216;where&#8217; non-V2 is far more frequent than V2 in Northern Norwegian, and that the picture gradually shifts to the opposite when one moves through Central Norwegian and Western Norwegian to Eastern Norwegian. The figures they provide can be summarised as in Table <xref ref-type="table" rid="T2">2</xref>.<xref ref-type="fn" rid="n2">2</xref></p>
<table-wrap id="T2">
<label>Table 2</label>
<caption>
<p>The balance between V2 and non-V2 in matrix <italic>wh</italic>-questions with short <italic>wh</italic>-constituents in the Norwegian part of the Nordic Dialect Corpus.</p>
</caption>
<table>
<tr>
<th valign="middle" align="left" rowspan="3" style="background-color:#f3f3f4;"></th>
<th valign="middle" align="center" colspan="2" style="background-color:#f3f3f4;">[what/who/where] + V2</th>
<th valign="middle" align="center" colspan="2" style="background-color:#f3f3f4;">[what/who/where] + non-V2</th>
</tr>
<tr>
<th colspan="4"><hr/></th>
</tr>
<tr>
<th valign="middle" align="center" style="background-color:#f3f3f4;">n</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">%</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">n</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">%</th>
</tr>
<tr>
<td colspan="5"><hr/></td>
</tr>
<tr>
<td valign="top" align="left">Northern Norwegian</td>
<td valign="top" align="right">81</td>
<td valign="top" align="right">21.3%</td>
<td valign="top" align="right">299</td>
<td valign="top" align="right">78.7%</td>
</tr>
<tr>
<td valign="top" align="left">Central Norwegian</td>
<td valign="top" align="right">50</td>
<td valign="top" align="right">43.1%</td>
<td valign="top" align="right">66</td>
<td valign="top" align="right">56.9%</td>
</tr>
<tr>
<td valign="top" align="left">Western Norwegian</td>
<td valign="top" align="right">124</td>
<td valign="top" align="right">56.6%</td>
<td valign="top" align="right">95</td>
<td valign="top" align="right">43.4%</td>
</tr>
<tr>
<td valign="top" align="left">Eastern Norwegian</td>
<td valign="top" align="right">147</td>
<td valign="top" align="right">79%</td>
<td valign="top" align="right">39</td>
<td valign="top" align="right">21%</td>
</tr>
<tr>
<td valign="top" align="left">Whole country</td>
<td valign="top" align="right">402</td>
<td valign="top" align="right">44.6%</td>
<td valign="top" align="right">499</td>
<td valign="top" align="right">55.4%</td>
</tr>
</table>
</table-wrap>
<p>On this issue the corpus data complement the questionnaire data in NSD as the latter only provide information about the acceptance of non-V2: the informants were never asked to judge matrix <italic>wh</italic>-questions with V2. Although we therefore do not know the relative preference of V2 versus non-V2, the production data from the corpus suggest that non-V2 is the unmarked option in Northern Norwegian dialects for the three short items &#8216;what&#8217;, &#8216;who&#8217; and &#8216;where&#8217; and that there is a gradual shift in preference as we move south.</p>
<p>Concerning the second observation, there exist both short and long variants for &#8216;when&#8217; in Norwegian dialects. Some dialects use the monosyllabic variant <italic>n&#229;r</italic>, which is the one used in the standard varieties, but the complex <italic>hva tid</italic>, literally &#8216;what time&#8217; is widespread, and the variants <italic>n&#229;r tid</italic> &#8216;when time&#8217; and <italic>hvor tid</italic> &#8216;where time&#8217; are also found. We should therefore consider what variants are used in the 21 instances of &#8216;when&#8217;-questions with non-V2. In this case we are also in the fortunate situation that one of the four <italic>wh</italic>-questions probing non-V2 in NSD is a &#8216;when&#8217;-clause (see above), thus allowing us to compare production and judgments directly at the level of the individual.</p>
<p>In Table <xref ref-type="table" rid="T3">3</xref> each informant is listed with whichever &#8216;when&#8217;-variant they used and how they judged the NSD &#8216;when&#8217;-question (#33 in the questionnaire). As we see, there are five instances with the short, monosyllabic item <italic>n&#229;r</italic> in non-V2 matrix <italic>wh</italic>-questions, produced by five different informants from three different locations in Central Norway. The informants&#8217; judgments of the NSD test sentence vary, but crucially the complex variant <italic>hva tid</italic> (adjusted for local pronunciation) was used during data collection also in this area, and the lesser acceptance of the test sentence may be due to the fact that the <italic>wh-</italic>expression used in the test is not the short variant the informants spontaneously use themselves.</p>
<table-wrap id="T3">
<label>Table 3</label>
<caption>
<p>Non-V2 &#8216;when&#8217;-question in the Nordic Syntax Database.</p>
</caption>
<table>
<tr>
<th valign="middle" align="left" style="background-color:#f3f3f4;">Region</th>
<th valign="middle" align="left" style="background-color:#f3f3f4;">Informant code</th>
<th valign="middle" align="left" style="background-color:#f3f3f4;">used by informant</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">NSD #33 score</th>
<th valign="middle" align="center" style="background-color:#f3f3f4;">n</th>
</tr>
<tr>
<td colspan="5"><hr/></td>
</tr>
<tr>
<td valign="top" align="left" rowspan="5">Central N<break/>(Tr&#248;ndelag, Nordm&#248;re)</td>
<td valign="top" align="left">inderoey_01um (young male)</td>
<td valign="top" align="left"><italic>n&#229;r &#8230;</italic></td>
<td valign="top" align="right">3</td>
<td valign="top" align="right">1</td>
</tr>
<tr>
<td valign="top" align="left">inderoey_02uk (young female)</td>
<td valign="top" align="left"><italic>n&#229;r &#8230;</italic></td>
<td valign="top" align="right">4</td>
<td valign="top" align="right">1</td>
</tr>
<tr>
<td valign="top" align="left">oppdal_10 (young female)</td>
<td valign="top" align="left"><italic>n&#229;r &#8230;</italic></td>
<td valign="top" align="right">1</td>
<td valign="top" align="right">1</td>
</tr>
<tr>
<td valign="top" align="left">surnadal_29 (young female)</td>
<td valign="top" align="left"><italic>n&#229;r &#8230;</italic></td>
<td valign="top" align="right">2</td>
<td valign="top" align="right">1</td>
</tr>
<tr>
<td valign="top" align="left">surnadal_28 (young female)</td>
<td valign="top" align="left"><italic>n&#229;r &#8230;</italic></td>
<td valign="top" align="right">2</td>
<td valign="top" align="right">1</td>
</tr>
<tr>
<td colspan="5"><hr/></td>
</tr>
<tr>
<td valign="top" align="left" rowspan="3">Northwestern N<break/>(Sunnm&#248;re, Nordfjord)</td>
<td valign="top" align="left">heroeyMR_03gm (older male)</td>
<td valign="top" align="left"><italic>ka ti &#8230;</italic></td>
<td valign="top" align="right">5</td>
<td valign="top" align="right">4</td>
</tr>
<tr>
<td valign="top" align="left">stryn_01um (young male)</td>
<td valign="top" align="left"><italic>ka ti &#8230;</italic></td>
<td valign="top" align="right">5</td>
<td valign="top" align="right">4</td>
</tr>
<tr>
<td valign="top" align="left">stryn_02uk (young female)</td>
<td valign="top" align="left"><italic>ka ti &#8230;</italic></td>
<td valign="top" align="right">5</td>
<td valign="top" align="right">6</td>
</tr>
<tr>
<td colspan="5"><hr/></td>
</tr>
<tr>
<td valign="top" align="left" rowspan="2">Southwestern N<break/>(Hordaland, Rogaland)</td>
<td valign="top" align="left">hjelmeland_01um (young male)</td>
<td valign="top" align="left"><italic>ka ti &#8230;</italic></td>
<td valign="top" align="right">5</td>
<td valign="top" align="right">1</td>
</tr>
<tr>
<td valign="top" align="left">bergen_02uk (young female)</td>
<td valign="top" align="left"><italic>korr ti &#8230;</italic></td>
<td valign="top" align="right">1</td>
<td valign="top" align="right">1</td>
</tr>
</table>
</table-wrap>
<p>There are, furthermore, 14 examples produced by three informants from two locations in Northwestern Norway, all of whom give the NSD test sentence the highest score. This is an area known for allowing complex <italic>wh-</italic>items with non-V2, and also here the production and judgment data are in harmony. The two remaining examples both involve complex <italic>wh</italic>-expressions. They are uttered by two informants from two places in Southwestern Norway: Hjelmeland and Bergen. The Hjelmeland informant gives the test sentence a high score, whereas the Bergen informant gives it a low score.</p>
<p>Of the 21 &#8216;when&#8217;-clauses there is therefore only one case where there is a clear incompatibility between production and judgment. The seemingly intermediate position of &#8216;when&#8217;-clauses is thus partly due to the fact that some of them involve a short, monosyllabic variant of the <italic>wh</italic>-expression and partly to the fact that only three informants produced most of the non-V2 cases (14 of 21).</p>
<p>Also for the manner <italic>how</italic> questions with non-V2 found in the corpus the simple~complex issue plays a role. The form of manner &#8216;how&#8217; varies to a considerable extent across Norwegian dialects (see Vangsnes 2008), and Vangsnes &amp; Westergaard (<xref ref-type="bibr" rid="B34">2014: 145</xref>) report that eight of the nine examples of non-V2 involves monosyllabic variants (<italic>korr, koss, h&#248;ss</italic>). Only the ninth example has a disyllabic variant (<italic>kelles</italic>).</p>
<p>In other words, the number of non-V2 questions with complex <italic>wh-</italic>expressions is even lower than it seems at first sight in Table <xref ref-type="table" rid="T3">3</xref>. The single &#8216;why&#8217;-clause involves a disyllabic <italic>wh</italic>-item (as is always the case in Norwegian dialects),<xref ref-type="fn" rid="n3">3</xref> and if we put together the &#8216;when&#8217; (21), &#8216;(manner) how&#8217; (8), &#8216;why&#8217; (1), and <italic>wh-</italic>phrases (9) &#8211; which amount to 39 &#8211; and subtract the ones with monosyllabic <italic>wh-</italic>expressions (5 <italic>when</italic> and 7 <italic>how</italic>), we are left with 27 non-V2 matrix questions with a complex <italic>wh-</italic>expression out of a total of 539, in other words 5%.</p>
<p>Map <xref ref-type="fig" rid="M6">6</xref> indicates the locations of all of the 27 examples, including the six mismatches (see <xref ref-type="bibr" rid="B34">Vangsnes &amp; Westergaard 2014: 145ff</xref>, for details).</p>
<fig id="M6">
<label>Map 6</label>
<caption>
<p>Locations at which the Nordic Dialect Corpus provides examples of non-V2 with complex wh-constituents.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65094/"/>
</fig>
<p>19 of the 27 examples are from four locations in Northwestern Norway, indicated by the light blue markers. Three of the cases are from three locations in Southwestern Norway, indicated by purple markers. Both of these areas are roughly the ones indicated by the letter A in Map <xref ref-type="fig" rid="M5">5</xref> above, hence where the NSD data suggests that informants by and large accept complex <italic>wh-</italic>phrases in matrix non-V2 questions.</p>
<p>The yellow icon marks the single example from a location in the northwest corner of the Eastern Norwegian dialect area, more specifically from the place Lom, which by the NSD data is part of the northwestern A-area: both of the test sentences with a complex <italic>wh</italic>-phrase (<italic>wh</italic>-subject and <italic>when</italic>) receive a high score at this location.</p>
<p>The remaining three cases run counter to the data in NSD. Two of them are uttered by informants at two locations in Central Norway, Oppdal and R&#248;ros, indicated by green markers, and at both locations the relevant test sentences receive a low score both in general and by the two specific individuals who uttered the corpus sentences in particular. The same holds for the final example from Bergen in Western Norway, indicated by a dark blue marker. Accordingly, these three examples constitute noise in the data that it would be worth following up in future studies of the topic.</p>
<p>Despite this slight discrepancy (3 out of 27 cases) and despite the low total number of complex <italic>wh-</italic>questions with non-V2, the overall picture we are left with when scrutinising the corpus data in the Nordic Dialect Corpus versus the judgment data in the Nordic Syntax Database is that there is a very good match between the two sources. Attempts at analysing these data from a more formal, generative perspective can be found in Westergaard et al. (<xref ref-type="bibr" rid="B43">2017</xref>). (See also Rognes <xref ref-type="bibr" rid="B25">2011</xref>, for a study focusing on one particular dialect area; and <xref ref-type="bibr" rid="B31">Vangsnes 2005</xref>; <xref ref-type="bibr" rid="B42">Westergaard &amp; Vangsnes 2005</xref>; for slightly older theoretical approaches.)</p>
</sec>
<sec>
<title>4.4 Subject wh-extraction and Nordg&#229;rd&#8217;s Generalisation</title>
<p>In the background section above we noted that Nordg&#229;rd (<xref ref-type="bibr" rid="B18">1985</xref>) established a correlation between dialects that allow non-V2 in matrix <italic>wh-</italic>questions and dialects that allow the insertion of the complementiser <italic>som</italic> before the trace of an extracted <italic>wh-</italic>subject. The questionnaire based data in the Nordic Syntax Database offers information on this issue.</p>
<p>As detailed in Bentzen (2014), eight test sentences for <italic>wh</italic>-extraction were also included in the questionnaire, probing both subject and object extraction, with and without the presence of either of the complementisers <italic>som</italic> and <italic>at</italic> &#8216;that&#8217; and also with a resumptive subject pronoun. The following three sentences tap into the Nordg&#229;rd&#8217;s generalisation (<xref ref-type="bibr" rid="B18">1985</xref>) and more broadly the so-called that-trace effect or COMP trace effect (see Pesetsky 2016 for an overview).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(18)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Hvem</p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p>tror</p></list-item>
<list-item><p>think</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>har</p></list-item>
<list-item><p>has</p></list-item>
</list>
<list list-type="word">
<list-item><p>gjort</p></list-item>
<list-item><p>done</p></list-item>
</list>
<list list-type="word">
<list-item><p>det?</p></list-item>
<list-item><p>it</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Hvem</p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p>tror</p></list-item>
<list-item><p>think</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>som</p></list-item>
<list-item><p><sc>SOM</sc></p></list-item>
</list>
<list list-type="word">
<list-item><p>har</p></list-item>
<list-item><p>has</p></list-item>
</list>
<list list-type="word">
<list-item><p>gjort</p></list-item>
<list-item><p>done</p></list-item>
</list>
<list list-type="word">
<list-item><p>det?</p></list-item>
<list-item><p>it</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>c.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Hvem</p></list-item>
<list-item><p>who</p></list-item>
</list>
<list list-type="word">
<list-item><p>tror</p></list-item>
<list-item><p>think</p></list-item>
</list>
<list list-type="word">
<list-item><p>du</p></list-item>
<list-item><p>you</p></list-item>
</list>
<list list-type="word">
<list-item><p>at</p></list-item>
<list-item><p>that</p></list-item>
</list>
<list list-type="word">
<list-item><p>har</p></list-item>
<list-item><p>has</p></list-item>
</list>
<list list-type="word">
<list-item><p>gjort</p></list-item>
<list-item><p>done</p></list-item>
</list>
<list list-type="word">
<list-item><p>det?</p></list-item>
<list-item><p>it</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="final-sentence">
<list-item><p>All: &#8216;Who do you think has done it?&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The following three maps (Map <xref ref-type="fig" rid="M7">7</xref>) details the acceptance of the three sentences in Central, Western and Eastern Norway, with white markers indicating high average scores, grey markers medium average scores, and black markers low average scores (cf. section 2).</p>
<fig id="M7">
<label>Map 7</label>
<caption>
<p>Average scores for (long) extraction of a <italic>wh</italic>-subject with: <bold>a)</bold> no complementiser in the embedded clause (18a); <bold>b)</bold> presence of the complementiser <italic>som</italic> in the embedded clause (18b); <bold>c)</bold> presence of the complementiser <italic>at</italic> in the embedded clause (18c).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65095/"/>
</fig>
<p>Map <xref ref-type="fig" rid="M7">7a</xref> (the leftmost one) clearly indicates that the sentence with no overt complementiser is accepted as good by everyone. Furthermore, the map in the middle shows that <italic>som</italic>-insertion mostly gets a high or medium score in Western and Central Norway but is largely rejected in Eastern Norway. The sentence with <italic>at-</italic>insertion is in contrast only fully accepted at some measure points in Eastern Norway with some medium scores further to the North in Central Norway.</p>
<p>The data visualised here partly support Nordg&#229;rd&#8217;s (<xref ref-type="bibr" rid="B18">1985</xref>) generalisation insofar that sentence (18b) is only accepted in those parts of the country where non-V2 is accepted. At the same time, it is also clear that far from all speakers who allow non-V2 allow <italic>som-</italic>insertion with extraction of a <italic>wh</italic>-subject. Still, the preference for no complementiser before a subject trace position over versions with an overt complementiser is a quite well-known fact from previous studies of Germanic languages, and it is also documented for object extraction (see <xref ref-type="bibr" rid="B2">Cowart 1997</xref>; <xref ref-type="bibr" rid="B6">Hawkins 2004</xref>; Bentzen 2014; <xref ref-type="bibr" rid="B27">Schippers 2017</xref>). On the presumption that the informants in the NSD survey were able to contrast the cases with and without complementiser when consulted, the overall lower acceptance for the COMP trace sentences should therefore come as no great surprise.</p>
<p>Furthermore, although the complementarity between <italic>som</italic>-insertion and <italic>at</italic>-insertion is not perfect given the many medium score locations in Central Norway in particular, Map <xref ref-type="fig" rid="M8">8</xref> shows that when we just compare measure points with a high score, the complementarity is quite clear: the grey dots mark locations with a high score for <italic>som</italic>-insertion and the blue ones a high score for <italic>at-</italic>insertion.</p>
<fig id="M8">
<label>Map 8</label>
<caption>
<p>Locations with high average scores for presence of complementiser <italic>som</italic> under subject wh-extraction (blue markers) versus locations with high average scores for presence of complementiser <italic>at</italic> under subject <italic>wh-</italic>extraction (grey markers).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/5133/file/65096/"/>
</fig>
<p>This map also shows that <italic>som</italic>-insertion is widely accepted in Northern Norway and that <italic>at-</italic>insertion is widely accepted in Finland Swedish.</p>
</sec>
<sec>
<title>4.5 Discussion</title>
<p>This exposition of how the issue of non-V2 in Norwegian matrix <italic>wh</italic>-questions was researched in the ScanDiaSyn project should have revealed some significant achievements and also some limitations. In many ways the data confirm what had already been established if one pieces together information from various sources in the existing literature, they establish this in a much more systematic and complete way. The questionnaire and corpus data furthermore also by and large confirm each other and thus strengthen the empirical basis.</p>
<p>One very clear limitation with the questionnaire data is that the number of test sentences is low. Relevant additional variables may not have been detected because of this. In particular, given the scarcity of non-V2 questions with complex <italic>wh-</italic>constituents in the corpus compared to the abundance of non-V2 with short <italic>wh</italic>-constituents, it would have been desirable to test out a broader range of complex <italic>wh-</italic>constituents so as to compare different kinds of <italic>wh-</italic>adverbials or <italic>wh</italic>-adverbials versus complex <italic>wh</italic>-arguments.</p>
<p>Furthermore, as pointed out above, there are quite clearly some cases of discrepancy between production and judgment data. A positive angle to that is that they identify areas and/or locations that need to be studied more carefully, and the northern part of Eastern Norway, marked as &#8220;?&#8221; in Map <xref ref-type="fig" rid="M5">5</xref>, in particular stands out as an area with an unclear pattern.</p>
<p>In this paper we have done little to put the data to statistical scrutiny. With data from over 100 locations and almost 400 individuals on a number of test sentences, the possibilities for doing so is certainly there, and one study which has approached the phenomenon in a systematic way by employing statistical methods is Westendorp (<xref ref-type="bibr" rid="B35">2017</xref>; <xref ref-type="bibr" rid="B36">2018</xref>). Without going into details, she argues that the statistics do not support all of the diachronic speculations put forth in Westergaard et al. (<xref ref-type="bibr" rid="B43">2017</xref>) as to how Norwegian dialects &#8211; as the only ones across North Germanic &#8211; have developed this particular violation of Verb Second. Westendorp&#8217;s statistical findings do however support the general idea that the phenomenon has started with short <italic>wh</italic>-constituents and later spread to questions with complex <italic>wh</italic>-expressions.</p>
</sec>
</sec>
<sec>
<title>5 Conclusion</title>
<p>In this paper we hope to have demonstrated the assets of having access to both a database of syntactic judgments and a searchable corpus of free speech when researching topics in dialect syntax. In the case of the two Nordic dialect infrastructures, the Nordic Syntax Database and the Nordic Dialect Corpus, the data have been collected from a well-distributed set of locations, and crucially both kinds of data have been collected from largely the same set of informants. The well-known shortcoming of a corpus that particular constructions or variables may be scarce is counterbalanced by the way a questionnaire can ensure information from all participants on the selected constructions/variables. On the other hand, the closed nature of a questionnaire, where topics must be decided beforehand can be contrasted to the more dynamic nature of a corpus in which data one had not thought of in advance may occur. Moreover, data from the two sources (judgment versus production) for the same individuals may confirm each other, but they may also be contradictory. The latter kind of situation may help to identify issues that need to be further investigated.</p>
<p>For the concrete topics that we have chosen to base our exposition on, we have seen that the data from dialect infrastructures may represent both a correction of received knowledge and a strengthened confirmation of existing knowledge. In the case of split infinitives both the judgment data and the corpus data clearly show that in spoken Norwegian placing a sentence adverb between the infinitival marker and the infinitive (split version) is much more preferred and used than placing it before the infinitival marker (unsplit version). This runs counter to the received knowledge that Norwegian allows both structures and in fact prefers the unsplit version: the data show that in Norwegian dialects the split version is both preferred and most commonly used, placing them more in line with Swedish than with Danish.</p>
<p>In the case of matrix <italic>wh-</italic>questions with non-V2 word order in Norwegian, the judgment data serve to confirm the rather complex pattern of variation that can be pieced together on the basis of the existing literature going back to the early 20<sup>th</sup> century. Furthermore, the corpus data serve to complement the questionnaire data insofar that the abundance of non-V2 <italic>wh-</italic>questions in all dialect regions but Eastern Norwegian confirms that the phenomenon is widespread, and we also see an increase in the use of it as we move northwards through the country. However, the abundance of examples applies just to questions with short <italic>wh-</italic>constituents: only 5% of the non-V2 <italic>wh-</italic>questions in the corpus contain complex <italic>wh</italic>-constituents, and they are also very few compared to V2 <italic>wh-</italic>questions with the same kind of constituents. To some extent this finding squares with the insight that complex non-V2 <italic>wh</italic>-questions are judged acceptable in two rather restricted areas (northwestern dialects and southwestern dialects), and almost all of the cases are indeed produced by speakers from these areas. But even within these restricted areas the complex non-V2 cases are fewer than complex V2 cases, and this entails that weight or complexity plays a role in production even in these dialects.</p>
<p>One clear shortcoming with the questionnaire data regarding the issue of <italic>wh</italic>-questions and V2, is that the number of test sentences were few. This is a direct effect of the topic being part of a general questionnaire that probed a wide range of topics. When administering data collection by questionnaires there is a limit to how much time one can keep the informants&#8217; attention and willingness to respond. Since the questionnaire and production data were collected at the same time by fieldworkers visiting the various locations this is a limitation that it is hard to come around unless one sets up ways of getting back to each individual informant at later points in time. In turn, administering a system for that is laborious and resource demanding and was not viable in the case of the ScanDiaSyn project.</p>
<p>In any event, it seems quite clear that the establishing of the Nordic research infrastructure for syntactic variation has moved the field of North Germanic dialect syntax several steps forward, and although others may be in a better position to judge it objectively, we also believe that the output of the research collaboration has served to revitalise the field of Nordic dialectology.</p>
</sec>
</body>
<back>
<fn-group>
<fn id="n1"><p>Vangsnes &amp; Westergaard (<xref ref-type="bibr" rid="B34">2014</xref>) report the number 22 for <italic>when-</italic>clauses with non-V2, but on closer examination it turns out that one of the hits were counted twice, and the number is adjusted accordingly here.</p></fn>
<fn id="n2"><p>Vangsnes &amp; Westergaard (op. cit.) separates out the district Agder, but we include it here in Western Norway to get the four main Norwegian dialect areas described in M&#230;hlum and R&#248;yneland (2012).</p></fn>
<fn id="n3"><p>The word for &#8216;why&#8217; in Norwegian dialects is always a complex item literally corresponding to &#8216;what-for&#8217; or &#8216;where-for&#8217;. The single &#8216;why&#8217;-example with non-V2 in the corpus has the variant <italic>koffor</italic>.</p></fn>
</fn-group>
<sec>
<title>Funding Information</title>
<p>The infrastructure was built with support from national and Nordic funding bodies: The Research Council of Norway, The Swedish Research Council, The Danish Research Council for Culture and Communication, and The Icelandic Research Fund, and the Nordic funding bodies NordForsk and NOS-HS which also provided important funding for network activities.</p>
</sec>
<sec>
<title>Competing Interests</title>
<p>The authors have no competing interests to declare.</p>
</sec>
<ref-list>
<ref id="B1"><label>1</label><mixed-citation publication-type="book"><string-name><surname>&#197;farli</surname>, <given-names>Tor A.</given-names></string-name> <year>1986</year>. <chapter-title>Some syntactic structures in a dialect of Norwegian</chapter-title>. <source>Working Papers in Linguistics</source> <volume>3</volume>. <fpage>93</fpage>&#8211;<lpage>111</lpage>. <publisher-loc>Trondheim</publisher-loc>: <publisher-name>University of Trondheim</publisher-name>.</mixed-citation></ref>
<ref id="B2"><label>2</label><mixed-citation publication-type="book"><string-name><surname>Cowart</surname>, <given-names>Wayne</given-names></string-name>. <year>1997</year>. <source>Experimental syntax: Applying objective methods to sentence judgments</source>. <publisher-loc>Thousand Oaks, CA</publisher-loc>: <publisher-name>SAGE Publications</publisher-name>.</mixed-citation></ref>
<ref id="B3"><label>3</label><mixed-citation publication-type="book"><string-name><surname>Elstad</surname>, <given-names>K&#229;re</given-names></string-name>. <year>1982</year>. <chapter-title>Nordnorske dialektar</chapter-title>. In <string-name><given-names>Tove</given-names> <surname>Bull</surname></string-name> &amp; <string-name><given-names>Kjellaug</given-names> <surname>Jetne</surname></string-name> (eds.), <source>Nordnorsk: Spr&#229;karv og spr&#229;kforhold i Nord-Noreg</source>, <fpage>11</fpage>&#8211;<lpage>100</lpage>. <publisher-loc>Oslo</publisher-loc>: <publisher-name>Det Norske Samlaget</publisher-name>.</mixed-citation></ref>
<ref id="B4"><label>4</label><mixed-citation publication-type="book"><string-name><surname>Faarlund</surname>, <given-names>Jan Terje</given-names></string-name>, <string-name><given-names>Svein</given-names> <surname>Lie</surname></string-name> &amp; <string-name><given-names>Kjell Ivar</given-names> <surname>Vannebo</surname></string-name>. <year>1997</year>. <source>Norsk referansegrammatikk</source>. <publisher-loc>Oslo</publisher-loc>: <publisher-name>Universitetsforlaget</publisher-name>.</mixed-citation></ref>
<ref id="B5"><label>5</label><mixed-citation publication-type="journal"><string-name><surname>Fiva</surname>, <given-names>Toril</given-names></string-name>. <year>1996</year>. <article-title>Sp&#248;rresetninger i Troms&#248;dialekten</article-title>. <source>Nordica Bergensia</source> <volume>9</volume>. <fpage>139</fpage>&#8211;<lpage>155</lpage>.</mixed-citation></ref>
<ref id="B6"><label>6</label><mixed-citation publication-type="book"><string-name><surname>Hawkins</surname>, <given-names>John A.</given-names></string-name> <year>2004</year>. <source>Efficiency and complexity in grammars</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1093/acprof:oso/9780199252695.001.0001</pub-id></mixed-citation></ref>
<ref id="B7"><label>7</label><mixed-citation publication-type="book"><string-name><surname>Huddleston</surname>, <given-names>Rodney D.</given-names></string-name> &amp; <string-name><given-names>Geoffrey K.</given-names> <surname>Pullum</surname></string-name>. <year>2002</year>. <source>The Cambridge grammar of the English language</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1017/9781316423530</pub-id></mixed-citation></ref>
<ref id="B8"><label>8</label><mixed-citation publication-type="book"><string-name><surname>Hulth&#233;n</surname>, <given-names>Lage</given-names></string-name>. <year>1947</year>. <source>Studier i j&#228;mf&#246;rande nunordisk syntax II</source>. <publisher-loc>G&#246;teborg</publisher-loc>: <publisher-name>Wettergren &amp; Kerbers F&#246;rlag</publisher-name>.</mixed-citation></ref>
<ref id="B9"><label>9</label><mixed-citation publication-type="book"><string-name><surname>Iversen</surname>, <given-names>Ragnvald</given-names></string-name>. <year>1918</year>. <source>Syntaksen i Troms&#248;bymaal</source>. <publisher-loc>Kristiania</publisher-loc>: <publisher-name>Bymaalslagets forlag</publisher-name>.</mixed-citation></ref>
<ref id="B10"><label>10</label><mixed-citation publication-type="book"><string-name><surname>Johannessen</surname>, <given-names>Janne Bondi</given-names></string-name>. <year>2017</year>. <chapter-title>Annotations in the Nordic Dialect Corpus</chapter-title>. In <string-name><given-names>Nancy</given-names> <surname>Ide</surname></string-name> &amp; <string-name><given-names>James</given-names> <surname>Pustejovsky</surname></string-name> (eds.), <source>Handbook of linguistic annotation</source>, Chapter 49. <publisher-name>Springer</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1007/978-94-024-0881-2_50</pub-id></mixed-citation></ref>
<ref id="B11"><label>11</label><mixed-citation publication-type="confproc"><string-name><surname>Johannessen</surname>, <given-names>Janne Bondi</given-names></string-name>, <string-name><given-names>Joel</given-names> <surname>Priestley</surname></string-name>, <string-name><given-names>Kristin</given-names> <surname>Hagen</surname></string-name>, <string-name><given-names>Tor Anders</given-names> <surname>&#197;farli</surname></string-name> &amp; <string-name><given-names>&#216;ystein Alexander</given-names> <surname>Vangsnes</surname></string-name>. <year>2009</year>. <article-title>The Nordic Dialect Corpus &#8211; An advanced research tool</article-title>. In <string-name><given-names>Kristiina</given-names> <surname>Jokinen</surname></string-name> &amp; <string-name><given-names>Eckhard</given-names> <surname>Bick</surname></string-name> (eds.), <conf-name>Proceedings of the 17th Nordic Conference of Computational Linguistics NODALIDA 2009</conf-name> (NEALT Proceedings Series Volume <volume>4</volume>), <fpage>73</fpage>&#8211;<lpage>80</lpage>. <conf-sponsor>Northern European Association for Language Technology (NEALT)</conf-sponsor>. Electronically published at Tartu University Library (Estonia). <uri>http://hdl.handle.net/10062/9206</uri>.</mixed-citation></ref>
<ref id="B12"><label>12</label><mixed-citation publication-type="confproc"><string-name><surname>Johannessen</surname>, <given-names>Janne Bondi</given-names></string-name>, <string-name><given-names>Lars</given-names> <surname>Nygaard</surname></string-name>, <string-name><given-names>Joel</given-names> <surname>Priestley</surname></string-name> &amp; <string-name><given-names>Anders</given-names> <surname>N&#248;klestad</surname></string-name>. <year>2008</year>. <article-title>Glossa: A multilingual, multimodal, configurable user interface</article-title>. In <string-name><given-names>Nicoletta</given-names> <surname>Calzolari</surname></string-name>, <string-name><given-names>Khalid</given-names> <surname>Choukri</surname></string-name>, <string-name><given-names>Bente</given-names> <surname>Maegaard</surname></string-name>, <string-name><given-names>Joseph</given-names> <surname>Mariani</surname></string-name>, <string-name><given-names>Jan</given-names> <surname>Odijk</surname></string-name>, <string-name><given-names>Stelios</given-names> <surname>Piperidis</surname></string-name> &amp; <string-name><given-names>Daniel</given-names> <surname>Tapias</surname></string-name> (eds.), <conf-name>Proceedings of the Sixth International Language Resources and Evaluation (LREC&#8217;08)</conf-name>, <fpage>617</fpage>&#8211;<lpage>621</lpage>. <conf-loc>Paris</conf-loc>: <conf-sponsor>European Language Resources Association (ELRA)</conf-sponsor>.</mixed-citation></ref>
<ref id="B13"><label>13</label><mixed-citation publication-type="book"><string-name><surname>Johannessen</surname>, <given-names>Janne Bondi</given-names></string-name>, <string-name><given-names>&#216;ystein Alexander</given-names> <surname>Vangsnes</surname></string-name>, <string-name><given-names>Joel</given-names> <surname>Priestley</surname></string-name> &amp; <string-name><given-names>Kristin</given-names> <surname>Hagen</surname></string-name>. <year>2014</year>. <chapter-title>A multilingual speech corpus of North-Germanic languages</chapter-title>. In <string-name><given-names>Tommaso</given-names> <surname>Raso</surname></string-name> &amp; <string-name><given-names>Heliana</given-names> <surname>Mello</surname></string-name> (eds.), <source>Spoken corpora and linguistic studies</source>, <fpage>69</fpage>&#8211;<lpage>83</lpage>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins Publishing Company</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1075/scl.61.02joh</pub-id></mixed-citation></ref>
<ref id="B14"><label>14</label><mixed-citation publication-type="journal"><string-name><surname>Lie</surname>, <given-names>Svein</given-names></string-name>. <year>1992</year>. <article-title>Ka du sei?</article-title> <source>Maal og Minne</source>, <fpage>62</fpage>&#8211;<lpage>77</lpage>.</mixed-citation></ref>
<ref id="B15"><label>15</label><mixed-citation publication-type="confproc"><string-name><surname>Lindstad</surname>, <given-names>Arne Martinus</given-names></string-name>, <string-name><given-names>Anders</given-names> <surname>N&#248;klestad</surname></string-name>, <string-name><given-names>Janne Bondi</given-names> <surname>Johannessen</surname></string-name> &amp; <string-name><given-names>&#216;ystein Alexander</given-names> <surname>Vangsnes</surname></string-name>. <year>2009</year>. <article-title>The Nordic Dialect Database: Mapping microsyntactic variation in the Scandinavian languages</article-title>. In <string-name><given-names>Kristiina</given-names> <surname>Jokinen</surname></string-name> &amp; <string-name><given-names>Eckhard</given-names> <surname>Bick</surname></string-name> (eds.), <conf-name>Proceedings of the 17th Nordic Conference of Computational Linguistics NODALIDA 2009</conf-name> (NEALT Proceedings Series Volume <volume>4</volume>), <fpage>283</fpage>&#8211;<lpage>286</lpage>. <conf-sponsor>Northern European Association for Language Technology (NEALT)</conf-sponsor>. Electronically published at Tartu University Library (Estonia). <uri>http://hdl.handle.net/10062/9206</uri>.</mixed-citation></ref>
<ref id="B16"><label>16</label><mixed-citation publication-type="book"><string-name><surname>Lohndal</surname>, <given-names>Terje</given-names></string-name>, <string-name><given-names>Marit</given-names> <surname>Westergaard</surname></string-name> &amp; <string-name><given-names>&#216;ystein A.</given-names> <surname>Vangsnes</surname></string-name>. Forthcoming. <chapter-title>Verb Second in Norwegian: Variation and acquisition</chapter-title>. In <string-name><given-names>Rebecca</given-names> <surname>Woods</surname></string-name>, <string-name><given-names>Sam</given-names> <surname>Wolfe</surname></string-name> &amp; <string-name><given-names>Theresa</given-names> <surname>Biberauer</surname></string-name> (eds.), <source>Rethinking verb second</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</mixed-citation></ref>
<ref id="B17"><label>17</label><mixed-citation publication-type="thesis"><string-name><surname>Nilsen</surname>, <given-names>Hilde</given-names></string-name>. <year>1996</year>. <source>Koff&#248;r dem sir det?</source> <publisher-loc>Troms&#248;</publisher-loc>: <publisher-name>University of Troms&#248;: Cand.philol</publisher-name>. thesis.</mixed-citation></ref>
<ref id="B18"><label>18</label><mixed-citation publication-type="thesis"><string-name><surname>Nordg&#229;rd</surname>, <given-names>Torbj&#248;rn</given-names></string-name>. <year>1985</year>. <source>Word order, binding and the empty category principle</source>. <publisher-loc>Trondheim</publisher-loc>: <publisher-name>University of Trondheim cand</publisher-name>. philol. thesis.</mixed-citation></ref>
<ref id="B19"><label>19</label><mixed-citation publication-type="webpage"><source>Nordic Atlas of Language Structures Online Journal (NALS)</source> <fpage>1</fpage>. <uri>https://journals.uio.no/index.php/NALS</uri>.</mixed-citation></ref>
<ref id="B20"><label>20</label><mixed-citation publication-type="webpage"><source>Nordic Atlas of Language Structures Online Journal (NALS)</source>, <fpage>2</fpage> [old site with thematic structure]. <uri>http://www.tekstlab.uio.no/nals#/project_info</uri>.</mixed-citation></ref>
<ref id="B21"><label>21</label><mixed-citation publication-type="webpage"><source>Nordic Dialect Corpus</source>. <uri>http://www.tekstlab.uio.no/nota/scandiasyn/</uri>.</mixed-citation></ref>
<ref id="B22"><label>22</label><mixed-citation publication-type="webpage"><source>Nordic Syntax Database</source>. <uri>http://www.tekstlab.uio.no/nota/scandiasyn/</uri>.</mixed-citation></ref>
<ref id="B23"><label>23</label><mixed-citation publication-type="book"><string-name><surname>Pedersen</surname>, <given-names>Karen Margrethe</given-names></string-name>. <year>2017</year>. <chapter-title>Syntaktiske oplysninger i dialektordb&#248;ger og store nationale ordb&#248;ger</chapter-title>. In <string-name><given-names>Jan-Ola</given-names> <surname>&#214;stman</surname></string-name>, et al. (eds.), <source>Ideologi, identitet, intervention: Nordisk dialektologi</source> <volume>10</volume>. <publisher-loc>Helsinki</publisher-loc>: <publisher-name>Nordica, University of Helsinki</publisher-name>.</mixed-citation></ref>
<ref id="B24"><label>24</label><mixed-citation publication-type="thesis"><string-name><surname>Reite</surname>, <given-names>Andr&#233;</given-names></string-name>. <year>2011</year>. <source>Sp&#248;rjing i skedsmokorsm&#97;&#778;let: Unders&#248;king og analyse av leddstelling og sp&#248;rjeord i interrogative hovudsetningar i talem&#97;&#778;let p&#97;&#778; Skedsmokorset</source>. <publisher-loc>Trondheim</publisher-loc>: <publisher-name>Norwegian University of Science and Technology (NTNU)</publisher-name> MA thesis.</mixed-citation></ref>
<ref id="B25"><label>25</label><mixed-citation publication-type="thesis"><string-name><surname>Rognes</surname>, <given-names>Stig</given-names></string-name>. <year>2011</year>. <source>V2, V3, V4 (and maybe even more): The syntax of questions in the Rogaland dialects of Norway</source>. <publisher-loc>Oslo</publisher-loc>: <publisher-name>University of Oslo</publisher-name> MA thesis.</mixed-citation></ref>
<ref id="B26"><label>26</label><mixed-citation publication-type="webpage"><source>ScanDiaSyn network</source>. <uri>http://websim.arkivert.uit.no/scandiasyn/index.html%3fLanguage=en</uri>.</mixed-citation></ref>
<ref id="B27"><label>27</label><mixed-citation publication-type="confproc"><string-name><surname>Schippers</surname>, <given-names>Ankelien</given-names></string-name>. <year>2017</year>. <article-title>On the variability of COMP-trace effects: A processing explanation</article-title>. <conf-name>Paper presented at Comparative Germanic Syntax Workshop</conf-name> <volume>32</volume>. <conf-loc>Trondheim</conf-loc>.</mixed-citation></ref>
<ref id="B28"><label>28</label><mixed-citation publication-type="thesis"><string-name><surname>Stroh-Wollin</surname>, <given-names>Ulla</given-names></string-name>. <year>2002</year>. <chapter-title><italic>Som</italic>-satser med och utan <italic>som</italic> [<italic>Som</italic>-sentences with and without <italic>som</italic>]</chapter-title>. <publisher-loc>Uppsala</publisher-loc>: <publisher-name>Uppsala University</publisher-name> doctoral dissertation.</mixed-citation></ref>
<ref id="B29"><label>29</label><mixed-citation publication-type="book"><string-name><surname>Taraldsen</surname>, <given-names>Knut Tarald</given-names></string-name>. <year>1986</year>. <chapter-title><italic>Som</italic> and the Binding Theory</chapter-title>. In <string-name><given-names>Lars</given-names> <surname>Hellan</surname></string-name> &amp; <string-name><given-names>Kirsti Koch</given-names> <surname>Christensen</surname></string-name> (eds.), <source>Topics in Scandinavian syntax</source>, <fpage>149</fpage>&#8211;<lpage>184</lpage>. <publisher-loc>Dordrecht</publisher-loc>: <publisher-name>Reidel</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1007/978-94-009-4572-2_8</pub-id></mixed-citation></ref>
<ref id="B30"><label>30</label><mixed-citation publication-type="book"><string-name><surname>Vangsnes</surname>, <given-names>&#216;ystein A.</given-names></string-name> <year>2004</year>. <chapter-title>On <italic>wh</italic>-questions and V2 across Norwegian dialects: A survey and some speculations</chapter-title>. <source>Working Papers in Scandinavian Syntax</source> <volume>73</volume>. <fpage>1</fpage>&#8211;<lpage>59</lpage>. <publisher-loc>Lund</publisher-loc>: <publisher-name>University of Lund</publisher-name>.</mixed-citation></ref>
<ref id="B31"><label>31</label><mixed-citation publication-type="book"><string-name><surname>Vangsnes</surname>, <given-names>&#216;ystein A.</given-names></string-name> <year>2005</year>. <chapter-title>Microparameters for Norwegian <italic>wh</italic>-grammars</chapter-title>. <source>Linguistic Variation Yearbook</source> <volume>5</volume>. <fpage>187</fpage>&#8211;<lpage>226</lpage>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>.</mixed-citation></ref>
<ref id="B32"><label>32</label><mixed-citation publication-type="journal"><string-name><surname>Vangsnes</surname>, <given-names>&#216;ystein A.</given-names></string-name> <year>2007a</year>. <article-title>Scandinavian Dialect Syntax (before and after) 2005</article-title>. <source>Nordlyd &#8211; Troms&#248; University Working Papers in Language &amp; Linguistics</source>, <fpage>7</fpage>&#8211;<lpage>24</lpage>.</mixed-citation></ref>
<ref id="B33"><label>33</label><mixed-citation publication-type="book"><string-name><surname>Vangsnes</surname>, <given-names>&#216;ystein A.</given-names></string-name> <year>2007b</year>. <chapter-title>ScanDiaSyn: Prosjektparaplyen Nordisk dialektsyntaks</chapter-title>. In <string-name><given-names>T.</given-names> <surname>Arboe</surname></string-name> (ed.), <source>Nordisk dialektologi og sociolingvistik</source>, <fpage>54</fpage>&#8211;<lpage>72</lpage>. <publisher-loc>&#197;rhus</publisher-loc>: <publisher-name>Peter Skautrup Centeret for Jysk Dialektforskning, &#197;rhus Universitet</publisher-name>.</mixed-citation></ref>
<ref id="B34"><label>34</label><mixed-citation publication-type="book"><string-name><surname>Vangsnes</surname>, <given-names>&#216;ystein A.</given-names></string-name> &amp; <string-name><given-names>Marit</given-names> <surname>Westergaard</surname></string-name>. <year>2014</year>. <chapter-title><italic>Ka korpuse fort&#230;ll?</italic> Om ordstilling i <italic>hv</italic>-sp&#248;rsm&#229;l i norske dialekter</chapter-title>. In <string-name><given-names>Janne Bondi</given-names> <surname>Johannessen</surname></string-name> &amp; <string-name><given-names>Kristin</given-names> <surname>Hagen</surname></string-name> (eds.), <source>Spr&#229;k i Norge og nabolanda. Ny forskning om talespr&#229;k</source>, <fpage>133</fpage>&#8211;<lpage>151</lpage>. <publisher-loc>Oslo</publisher-loc>: <publisher-name>Novus</publisher-name>.</mixed-citation></ref>
<ref id="B35"><label>35</label><mixed-citation publication-type="thesis"><string-name><surname>Westendorp</surname>, <given-names>Maud</given-names></string-name>. <year>2017</year>. <source>Development and variation of non-V2 order in Norwegian wh-questions</source>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>University of Amsterdam MA</publisher-name> thesis.</mixed-citation></ref>
<ref id="B36"><label>36</label><mixed-citation publication-type="journal"><string-name><surname>Westendorp</surname>, <given-names>Maud</given-names></string-name>. <year>2018</year>. <article-title>New methodologies in the Nordic Syntax Database: Word order variation in Norwegian <italic>wh</italic>-questions</article-title>. In <source>Nordic Atlas of Language Structures Online Journal</source>, <volume>3</volume>(<issue>1</issue>). <fpage>1</fpage>&#8211;<lpage>18</lpage>.</mixed-citation></ref>
<ref id="B37"><label>37</label><mixed-citation publication-type="journal"><string-name><surname>Westergaard</surname>, <given-names>Marit</given-names></string-name>. <year>2003</year>. <article-title>Word order in <italic>wh</italic>-questions in a North Norwegian dialect: Some evidence from an acquisition study</article-title>. <source>Nordic Journal of Linguistics</source> <volume>26</volume>. <fpage>81</fpage>&#8211;<lpage>109</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0332586503001021</pub-id></mixed-citation></ref>
<ref id="B38"><label>38</label><mixed-citation publication-type="journal"><string-name><surname>Westergaard</surname>, <given-names>Marit</given-names></string-name>. <year>2005</year>. <article-title>Optional word order in <italic>wh</italic>-questions in two Norwegian dialects: A diachronic analysis of synchronic variation</article-title>. <source>Nordic Journal of Linguistics</source> <volume>28</volume>. <fpage>269</fpage>&#8211;<lpage>296</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0332586505001459</pub-id></mixed-citation></ref>
<ref id="B39"><label>39</label><mixed-citation publication-type="journal"><string-name><surname>Westergaard</surname>, <given-names>Marit</given-names></string-name>. <year>2009a</year>. <article-title>Microvariation as diachrony: A view from acquisition</article-title>. <source>Journal of Comparative Germanic Linguistics</source> <volume>12</volume>. <fpage>49</fpage>&#8211;<lpage>79</lpage>. DOI: <pub-id pub-id-type="doi">10.1007/s10828-009-9025-9</pub-id></mixed-citation></ref>
<ref id="B40"><label>40</label><mixed-citation publication-type="book"><string-name><surname>Westergaard</surname>, <given-names>Marit</given-names></string-name>. <year>2009b</year>. <source>The acquisition of word order: Micro-cues, information structure, and economy</source>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1075/la.145</pub-id></mixed-citation></ref>
<ref id="B41"><label>41</label><mixed-citation publication-type="book"><string-name><surname>Westergaard</surname>, <given-names>Marit</given-names></string-name>. <year>2017</year>. <chapter-title>Word order and verb movement in Norwegian <italic>wh</italic>-questions: A comparison of production and judgment data</chapter-title>. In <string-name><given-names>Bettelou</given-names> <surname>Los</surname></string-name> &amp; <string-name><given-names>Pieter</given-names> <surname>de Haan</surname></string-name> (eds.), <source>Word order change in acquisition and contact: Essays in honour of Ans van Kemenade</source>, <fpage>35</fpage>&#8211;<lpage>56</lpage>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1075/la.243.03wes</pub-id></mixed-citation></ref>
<ref id="B42"><label>42</label><mixed-citation publication-type="journal"><string-name><surname>Westergaard</surname>, <given-names>Marit</given-names></string-name> &amp; <string-name><given-names>&#216;ystein A.</given-names> <surname>Vangsnes</surname></string-name>. <year>2005</year>. <article-title><italic>Wh</italic>-questions, V2, and the left periphery of three Norwegian dialects</article-title>. <source>Journal of Comparative Germanic Linguistics</source> <volume>8</volume>. <fpage>119</fpage>&#8211;<lpage>160</lpage>. DOI: <pub-id pub-id-type="doi">10.1007/s10828-004-0292-1</pub-id></mixed-citation></ref>
<ref id="B43"><label>43</label><mixed-citation publication-type="journal"><string-name><surname>Westergaard</surname>, <given-names>Marit</given-names></string-name>, <string-name><given-names>&#216;ystein A.</given-names> <surname>Vangsnes</surname></string-name> &amp; <string-name><given-names>Terje</given-names> <surname>Lohndal</surname></string-name>. <year>2017</year>. <article-title>Variation and change in Norwegian <italic>wh</italic>-questions: The role of the complementizer <italic>som</italic></article-title>. <source>Linguistic variation</source> <volume>17</volume>(<issue>1</issue>). <fpage>8</fpage>&#8211;<lpage>43</lpage>.</mixed-citation></ref>
</ref-list>
</back>
</article>