bibliographic newsletter 1 taylor: bibliographic newsletter published by cu scholar, 1975 2 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/1 colorado research in linguistics 5-1975 bibliographic newsletter allan r. taylor recommended citation tmp.1538170152.pdf.o70un locative expressions in siouan and caddoan 1 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 2 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/4 doi: https://doi.org/10.25810/h80h-gk74 3 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 4 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/4 doi: https://doi.org/10.25810/h80h-gk74 5 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 6 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/4 doi: https://doi.org/10.25810/h80h-gk74 7 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 8 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/4 doi: https://doi.org/10.25810/h80h-gk74 9 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 10 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/4 doi: https://doi.org/10.25810/h80h-gk74 11 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 12 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/4 doi: https://doi.org/10.25810/h80h-gk74 13 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 14 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/4 doi: https://doi.org/10.25810/h80h-gk74 15 rood: locative expressions in siouan and caddoan published by cu scholar, 1979 colorado research in linguistics 5-1979 locative expressions in siouan and caddoan david s. rood recommended citation tmp.1538088556.pdf.cykdt the bivium syndrome in the history of semiotics 1 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 2 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 3 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 4 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 5 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 6 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 7 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 8 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 9 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 10 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 11 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 12 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 13 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 14 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 15 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 16 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 17 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 18 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 19 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 20 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 21 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 22 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 23 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 24 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/3 25 romeo: the bivium syndrome in the history of semiotics published by cu scholar, 1977 colorado research in linguistics 5-1977 the bivium syndrome in the history of semiotics luigi romeo recommended citation tmp.1538168672.pdf.q1fii non-psychological deep structures? 1 rood: non-psychological deep structures? published by cu scholar, 1976 2 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/3 3 rood: non-psychological deep structures? published by cu scholar, 1976 4 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/3 5 rood: non-psychological deep structures? published by cu scholar, 1976 6 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/3 7 rood: non-psychological deep structures? published by cu scholar, 1976 colorado research in linguistics 5-1976 non-psychological deep structures? david rood recommended citation tmp.1538171026.pdf.z3o6i phonetics of kanji and possible psycholinguistic correlates: notes by a novice 1 m en n: p ho ne tic s o f k an ji an d po ss ib le p sy ch ol in gu ist ic c or re la te s: n ot es b y a n ov ic e pu bl ish ed b y c u s ch ol ar , 1 98 9 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ x3 xp -v g4 2 3 m en n: p ho ne tic s o f k an ji an d po ss ib le p sy ch ol in gu ist ic c or re la te s: n ot es b y a n ov ic e pu bl ish ed b y c u s ch ol ar , 1 98 9 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ x3 xp -v g4 2 5 m en n: p ho ne tic s o f k an ji an d po ss ib le p sy ch ol in gu ist ic c or re la te s: n ot es b y a n ov ic e pu bl ish ed b y c u s ch ol ar , 1 98 9 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ x3 xp -v g4 2 colorado research in linguistics 5-1989 phonetics of kanji and possible psycholinguistic correlates: notes by a novice lise menn recommended citation tmp.1538225926.pdf.wbaie preparing lakhota teaching materials: a progress report 1 rood and taylor: preparing lakhota teaching materials: a progress report published by cu scholar, 1973 2 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/4 doi: https://doi.org/10.25810/1wft-sh52 3 rood and taylor: preparing lakhota teaching materials: a progress report published by cu scholar, 1973 4 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/4 doi: https://doi.org/10.25810/1wft-sh52 5 rood and taylor: preparing lakhota teaching materials: a progress report published by cu scholar, 1973 6 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/4 doi: https://doi.org/10.25810/1wft-sh52 7 rood and taylor: preparing lakhota teaching materials: a progress report published by cu scholar, 1973 8 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/4 doi: https://doi.org/10.25810/1wft-sh52 9 rood and taylor: preparing lakhota teaching materials: a progress report published by cu scholar, 1973 10 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/4 doi: https://doi.org/10.25810/1wft-sh52 colorado research in linguistics 5-1973 preparing lakhota teaching materials: a progress report david s. rood allan r. taylor recommended citation tmp.1538173225.pdf.kd4n7 politeness and subjunctive in spanish and japanese 1 lo za no a nd t ak ah ar a: p ol ite ne ss a nd s ub ju nc tiv e in s pa ni sh a nd ja pa ne se pu bl ish ed b y c u s ch ol ar , 1 98 6 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8q qr -n p4 8 3 lo za no a nd t ak ah ar a: p ol ite ne ss a nd s ub ju nc tiv e in s pa ni sh a nd ja pa ne se pu bl ish ed b y c u s ch ol ar , 1 98 6 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8q qr -n p4 8 colorado research in linguistics 1986 politeness and subjunctive in spanish and japanese anthony lozano kumiko takahara recommended citation tmp.1538224176.pdf.gtp92 judaeo-arabic scholarship and sanctius' antecedents 1 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 2 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 3 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 4 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 5 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 6 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 7 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 8 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 9 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 10 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 11 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 12 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 13 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 14 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 15 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 16 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 17 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 18 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 19 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 20 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 21 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 22 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 23 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 24 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 25 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 26 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 27 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 28 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 29 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 30 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/2 31 breva-claramonte: judaeo-arabic scholarship and sanctius' antecedents published by cu scholar, 1977 colorado research in linguistics 5-1977 judaeo-arabic scholarship and sanctius' antecedents manuel breva-claramonte recommended citation tmp.1538168556.pdf.u22fw the object complement in bahasa malaysia 1 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 2 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 3 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 4 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 5 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 6 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 7 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 8 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 9 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 10 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 11 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 12 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 13 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 14 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 15 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 16 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 17 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 18 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 19 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 20 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 21 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 22 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 23 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 24 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 25 devillers: the object complement in bahasa malaysia published by cu scholar, 1974 26 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/1 doi: https://doi.org/10.25810/ty2y-dx29 colorado research in linguistics 5-1974 the object complement in bahasa malaysia colette devillers recommended citation tmp.1538172006.pdf.zqivk wichita: an unusual phonology system 1 rood: wichita: an unusual phonology system published by cu scholar, 1971 2 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 3 rood: wichita: an unusual phonology system published by cu scholar, 1971 4 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 5 rood: wichita: an unusual phonology system published by cu scholar, 1971 6 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 7 rood: wichita: an unusual phonology system published by cu scholar, 1971 8 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 9 rood: wichita: an unusual phonology system published by cu scholar, 1971 10 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 11 rood: wichita: an unusual phonology system published by cu scholar, 1971 12 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 13 rood: wichita: an unusual phonology system published by cu scholar, 1971 14 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 15 rood: wichita: an unusual phonology system published by cu scholar, 1971 16 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 17 rood: wichita: an unusual phonology system published by cu scholar, 1971 18 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 19 rood: wichita: an unusual phonology system published by cu scholar, 1971 20 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 21 rood: wichita: an unusual phonology system published by cu scholar, 1971 22 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 23 rood: wichita: an unusual phonology system published by cu scholar, 1971 24 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/2 doi: https://doi.org/10.25810/a3tf-4246 colorado research in linguistics 12-1971 wichita: an unusual phonology system david s. rood recommended citation tmp.1538174073.pdf.ybjfb a note on the historiography of sanskrit linguistics 1 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 2 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/4 doi: https://doi.org/10.25810/rsv9-cj41 3 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 4 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/4 doi: https://doi.org/10.25810/rsv9-cj41 5 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 6 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/4 doi: https://doi.org/10.25810/rsv9-cj41 7 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 8 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/4 doi: https://doi.org/10.25810/rsv9-cj41 9 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 10 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/4 doi: https://doi.org/10.25810/rsv9-cj41 11 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 12 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/4 doi: https://doi.org/10.25810/rsv9-cj41 13 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 14 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/4 doi: https://doi.org/10.25810/rsv9-cj41 15 romeo: a note on the historiography of sanskrit linguistics published by cu scholar, 1974 colorado research in linguistics 5-1974 a note on the historiography of sanskrit linguistics luigi romeo recommended citation tmp.1538172319.pdf.shnnz some lakhota presuppositions 1 rood: some lakhota presuppositions published by cu scholar, 1977 2 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/6 3 rood: some lakhota presuppositions published by cu scholar, 1977 4 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/6 5 rood: some lakhota presuppositions published by cu scholar, 1977 6 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/6 7 rood: some lakhota presuppositions published by cu scholar, 1977 8 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/6 9 rood: some lakhota presuppositions published by cu scholar, 1977 10 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/6 11 rood: some lakhota presuppositions published by cu scholar, 1977 12 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/6 13 rood: some lakhota presuppositions published by cu scholar, 1977 14 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/6 15 rood: some lakhota presuppositions published by cu scholar, 1977 colorado research in linguistics 5-1977 some lakhota presuppositions david rood recommended citation tmp.1538169263.pdf.k8yna three-term space deixis 1 ta ka ha ra : t hr ee -t er m s pa ce d ei xi s pu bl ish ed b y c u s ch ol ar , 1 98 9 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ m 94 sw w 46 3 ta ka ha ra : t hr ee -t er m s pa ce d ei xi s pu bl ish ed b y c u s ch ol ar , 1 98 9 colorado research in linguistics 5-1989 three-term space deixis kumiko takahara recommended citation tmp.1538226078.pdf.narom perception of rhythm in english and of nonspeech analogues 1 be ll an d fo w le r: pe rc ep tio n of r hy th m in e ng lis h an d of n on sp ee ch a na lo gu es pu bl ish ed b y c u s ch ol ar , 1 98 6 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ cj qv -9 y2 8 3 be ll an d fo w le r: pe rc ep tio n of r hy th m in e ng lis h an d of n on sp ee ch a na lo gu es pu bl ish ed b y c u s ch ol ar , 1 98 6 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ cj qv -9 y2 8 colorado research in linguistics 1986 perception of rhythm in english and of nonspeech analogues alan bell carol fowler recommended citation tmp.1538223671.pdf.zkzex structural analysis of the verb copying construction in mandarin chinese 1 li u: s tr uc tu ra l a na ly sis o f t he v er b c op yi ng c on st ru ct io n in m an da rin c hi ne se pu bl ish ed b y c u s ch ol ar , 1 98 9 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8k sv -m 58 3 3 li u: s tr uc tu ra l a na ly sis o f t he v er b c op yi ng c on st ru ct io n in m an da rin c hi ne se pu bl ish ed b y c u s ch ol ar , 1 98 9 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8k sv -m 58 3 5 li u: s tr uc tu ra l a na ly sis o f t he v er b c op yi ng c on st ru ct io n in m an da rin c hi ne se pu bl ish ed b y c u s ch ol ar , 1 98 9 colorado research in linguistics 5-1989 structural analysis of the verb copying construction in mandarin chinese mei-chun liu recommended citation tmp.1538225682.pdf.dpzk8 case, a deeper matter 1 takahara: case, a deeper matter published by cu scholar, 1971 2 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 3 takahara: case, a deeper matter published by cu scholar, 1971 4 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 5 takahara: case, a deeper matter published by cu scholar, 1971 6 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 7 takahara: case, a deeper matter published by cu scholar, 1971 8 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 9 takahara: case, a deeper matter published by cu scholar, 1971 10 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 11 takahara: case, a deeper matter published by cu scholar, 1971 12 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 13 takahara: case, a deeper matter published by cu scholar, 1971 14 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 15 takahara: case, a deeper matter published by cu scholar, 1971 16 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 17 takahara: case, a deeper matter published by cu scholar, 1971 18 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 19 takahara: case, a deeper matter published by cu scholar, 1971 20 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 21 takahara: case, a deeper matter published by cu scholar, 1971 22 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/3 doi: https://doi.org/10.25810/zn5x-ap78 colorado research in linguistics 12-1971 case, a deeper matter kumiko takahara recommended citation tmp.1538174149.pdf.17lke how abstract is abstract? 1 jensen: how abstract is abstract? published by cu scholar, 1974 2 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 3 jensen: how abstract is abstract? published by cu scholar, 1974 4 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 5 jensen: how abstract is abstract? published by cu scholar, 1974 6 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 7 jensen: how abstract is abstract? published by cu scholar, 1974 8 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 9 jensen: how abstract is abstract? published by cu scholar, 1974 10 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 11 jensen: how abstract is abstract? published by cu scholar, 1974 12 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 13 jensen: how abstract is abstract? published by cu scholar, 1974 14 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 15 jensen: how abstract is abstract? published by cu scholar, 1974 16 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 17 jensen: how abstract is abstract? published by cu scholar, 1974 18 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/2 doi: https://doi.org/10.25810/d0k2-km61 19 jensen: how abstract is abstract? published by cu scholar, 1974 colorado research in linguistics 5-1974 how abstract is abstract? john t. jensen recommended citation tmp.1538172093.pdf.9kwbb the articulatory syllable: saussure to stetson 1 be ll: t he a rt ic ul at or y sy lla bl e: s au ss ur e to s te ts on pu bl ish ed b y c u s ch ol ar , 1 98 6 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ pk 43 -c m 55 3 be ll: t he a rt ic ul at or y sy lla bl e: s au ss ur e to s te ts on pu bl ish ed b y c u s ch ol ar , 1 98 6 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ pk 43 -c m 55 5 be ll: t he a rt ic ul at or y sy lla bl e: s au ss ur e to s te ts on pu bl ish ed b y c u s ch ol ar , 1 98 6 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ pk 43 -c m 55 colorado research in linguistics 1986 the articulatory syllable: saussure to stetson alan bell recommended citation tmp.1538223778.pdf.jtit9 a european loanword of early date in eastern north america 1 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 2 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 3 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 4 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 5 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 6 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 7 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 8 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 9 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 10 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 11 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 12 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 13 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 14 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 15 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 16 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 17 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 18 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 19 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 20 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 21 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 22 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 23 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 24 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/5 doi: https://doi.org/10.25810/bf0d-4b87 25 taylor: a european loanword of early date in eastern north america published by cu scholar, 1973 colorado research in linguistics 5-1973 a european loanword of early date in eastern north america allan r. taylor recommended citation tmp.1538173330.pdf.jggrg 'to be' in russian 1 tuniks: 'to be' in russian published by cu scholar, 1972 2 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 3 tuniks: 'to be' in russian published by cu scholar, 1972 4 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 5 tuniks: 'to be' in russian published by cu scholar, 1972 6 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 7 tuniks: 'to be' in russian published by cu scholar, 1972 8 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 9 tuniks: 'to be' in russian published by cu scholar, 1972 10 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 11 tuniks: 'to be' in russian published by cu scholar, 1972 12 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 13 tuniks: 'to be' in russian published by cu scholar, 1972 14 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 15 tuniks: 'to be' in russian published by cu scholar, 1972 16 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 17 tuniks: 'to be' in russian published by cu scholar, 1972 18 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/3 doi: https://doi.org/10.25810/0qmx-tj15 19 tuniks: 'to be' in russian published by cu scholar, 1972 colorado research in linguistics 10-1972 'to be' in russian galina tuniks recommended citation tmp.1538173828.pdf.q3ety a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' a mandan text collected by edward kennard 1 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 2 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 3 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 4 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 5 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 6 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 7 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 8 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 9 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 10 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 11 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 12 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 13 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 14 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 15 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 16 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 17 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 18 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 19 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 20 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 21 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 22 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 23 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 24 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 25 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 26 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 27 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 28 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 29 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 30 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 31 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 32 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 33 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 34 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 35 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 36 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 37 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 38 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 39 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 40 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 41 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 42 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 43 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 44 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 45 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 46 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 47 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 48 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 49 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 50 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 51 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 52 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 53 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 54 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 55 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 56 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 57 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 58 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 59 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 60 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 61 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 62 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 63 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 64 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 65 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 66 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 67 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 68 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 69 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 70 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 71 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 72 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 73 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 74 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 75 coberly: a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' published by cu scholar, 1979 76 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/2 doi: https://doi.org/10.25810/6r6h-7n29 colorado research in linguistics 5-1979 a text analysis and brief grammatical sketch based on 'trickster challenges the buffalo' a mandan text collected by edward kennard mary coberly recommended citation tmp.1538088219.pdf.qzfi2 some aspects of complex sentence structure in bahasa malaysia 1 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 2 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 3 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 4 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 5 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 6 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 7 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 8 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 9 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 10 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 11 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 12 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 13 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 14 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 15 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 16 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/6 doi: https://doi.org/10.25810/zj41-7b85 17 valerga-aráoz: some aspects of complex sentence structure in bahasa malaysia published by cu scholar, 1974 colorado research in linguistics 5-1974 some aspects of complex sentence structure in bahasa malaysia maria m. valerga-aráoz recommended citation tmp.1538172533.pdf.gwwxh word-based morphology: some problems from a polysynthetic language 1 a xe lro d: w or dba se d m or ph ol og y: s om e pr ob le m s f ro m a p ol ys yn th et ic l an gu ag e pu bl ish ed b y c u s ch ol ar , 1 98 9 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ vd w 9r0 07 3 a xe lro d: w or dba se d m or ph ol og y: s om e pr ob le m s f ro m a p ol ys yn th et ic l an gu ag e pu bl ish ed b y c u s ch ol ar , 1 98 9 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ vd w 9r0 07 5 a xe lro d: w or dba se d m or ph ol og y: s om e pr ob le m s f ro m a p ol ys yn th et ic l an gu ag e pu bl ish ed b y c u s ch ol ar , 1 98 9 colorado research in linguistics 5-1989 word-based morphology: some problems from a polysynthetic language melissa axelrod recommended citation tmp.1538225596.pdf.z2h4d morphology and polysynthetic languages 1 ro od : m or ph ol og y an d po ly sy nt he tic l an gu ag es pu bl ish ed b y c u s ch ol ar , 1 98 9 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ em yd -y d0 6 3 ro od : m or ph ol og y an d po ly sy nt he tic l an gu ag es pu bl ish ed b y c u s ch ol ar , 1 98 9 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ em yd -y d0 6 5 ro od : m or ph ol og y an d po ly sy nt he tic l an gu ag es pu bl ish ed b y c u s ch ol ar , 1 98 9 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ em yd -y d0 6 colorado research in linguistics 5-1989 morphology and polysynthetic languages david s. rood recommended citation tmp.1538226011.pdf.x_wz6 a reanalysis of the biloxi causative 1 k oo nt z: a r ea na ly sis o f t he b ilo xi c au sa tiv e pu bl ish ed b y c u s ch ol ar , 1 98 6 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ fs 7b -r h5 3 3 k oo nt z: a r ea na ly sis o f t he b ilo xi c au sa tiv e pu bl ish ed b y c u s ch ol ar , 1 98 6 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ fs 7b -r h5 3 5 k oo nt z: a r ea na ly sis o f t he b ilo xi c au sa tiv e pu bl ish ed b y c u s ch ol ar , 1 98 6 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ fs 7b -r h5 3 colorado research in linguistics 1986 a reanalysis of the biloxi causative john koontz recommended citation tmp.1538223892.pdf.n9a4l some aragonese morphophonemics 1 tiberio: some aragonese morphophonemics published by cu scholar, 1972 2 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 3 tiberio: some aragonese morphophonemics published by cu scholar, 1972 4 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 5 tiberio: some aragonese morphophonemics published by cu scholar, 1972 6 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 7 tiberio: some aragonese morphophonemics published by cu scholar, 1972 8 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 9 tiberio: some aragonese morphophonemics published by cu scholar, 1972 10 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 11 tiberio: some aragonese morphophonemics published by cu scholar, 1972 12 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 13 tiberio: some aragonese morphophonemics published by cu scholar, 1972 14 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 15 tiberio: some aragonese morphophonemics published by cu scholar, 1972 16 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 17 tiberio: some aragonese morphophonemics published by cu scholar, 1972 18 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 19 tiberio: some aragonese morphophonemics published by cu scholar, 1972 20 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 21 tiberio: some aragonese morphophonemics published by cu scholar, 1972 22 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 23 tiberio: some aragonese morphophonemics published by cu scholar, 1972 24 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 25 tiberio: some aragonese morphophonemics published by cu scholar, 1972 26 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 27 tiberio: some aragonese morphophonemics published by cu scholar, 1972 28 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 29 tiberio: some aragonese morphophonemics published by cu scholar, 1972 30 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 31 tiberio: some aragonese morphophonemics published by cu scholar, 1972 32 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 33 tiberio: some aragonese morphophonemics published by cu scholar, 1972 34 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 35 tiberio: some aragonese morphophonemics published by cu scholar, 1972 36 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 37 tiberio: some aragonese morphophonemics published by cu scholar, 1972 38 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 39 tiberio: some aragonese morphophonemics published by cu scholar, 1972 40 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 41 tiberio: some aragonese morphophonemics published by cu scholar, 1972 42 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 43 tiberio: some aragonese morphophonemics published by cu scholar, 1972 44 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 45 tiberio: some aragonese morphophonemics published by cu scholar, 1972 46 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 47 tiberio: some aragonese morphophonemics published by cu scholar, 1972 48 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/2 doi: https://doi.org/10.25810/3azk-sp43 49 tiberio: some aragonese morphophonemics published by cu scholar, 1972 colorado research in linguistics 10-1972 some aragonese morphophonemics gaio e. tiberio recommended citation tmp.1538173747.pdf.1zjbj when is a grammatical disorder a disorder of grammar? 1 m en n: w he n is a g ra m m at ic al d iso rd er a d iso rd er o f g ra m m ar ? pu bl ish ed b y c u s ch ol ar , 1 98 6 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ hc 2y -7 h3 8 3 m en n: w he n is a g ra m m at ic al d iso rd er a d iso rd er o f g ra m m ar ? pu bl ish ed b y c u s ch ol ar , 1 98 6 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ hc 2y -7 h3 8 5 m en n: w he n is a g ra m m at ic al d iso rd er a d iso rd er o f g ra m m ar ? pu bl ish ed b y c u s ch ol ar , 1 98 6 colorado research in linguistics 1986 when is a grammatical disorder a disorder of grammar? lise menn recommended citation tmp.1538224300.pdf.fdbbg postpositions in awutu 1 frajzyngier: postpositions in awutu published by cu scholar, 1971 2 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 3 frajzyngier: postpositions in awutu published by cu scholar, 1971 4 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 5 frajzyngier: postpositions in awutu published by cu scholar, 1971 6 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 7 frajzyngier: postpositions in awutu published by cu scholar, 1971 8 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 9 frajzyngier: postpositions in awutu published by cu scholar, 1971 10 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 11 frajzyngier: postpositions in awutu published by cu scholar, 1971 12 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 13 frajzyngier: postpositions in awutu published by cu scholar, 1971 14 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 15 frajzyngier: postpositions in awutu published by cu scholar, 1971 16 colorado research in linguistics, vol. 1 [1971] https://scholar.colorado.edu/cril/vol1/iss1/1 doi: https://doi.org/10.25810/bvwq-1y91 17 frajzyngier: postpositions in awutu published by cu scholar, 1971 colorado research in linguistics 12-1971 postpositions in awutu zygmunt frajzyngier recommended citation tmp.1538173982.pdf.stcvh pragmatics in the semantic description of the verbs of 'giving' 1 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 2 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 3 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 4 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 5 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 6 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 7 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 8 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 9 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 10 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 11 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 12 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 13 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 14 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 15 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 16 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/5 17 takahara: pragmatics in the semantic description of the verbs of 'giving' published by cu scholar, 1976 colorado research in linguistics 5-1976 pragmatics in the semantic description of the verbs of 'giving' kumiko takahara recommended citation tmp.1538171447.pdf.z_pqf a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children 1 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 2 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 3 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 4 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 5 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 6 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 7 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 8 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 9 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 10 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 11 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 12 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 13 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 14 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 15 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 16 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/1 doi: https://doi.org/10.25810/9j3d-3j64 17 coberly: a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children published by cu scholar, 1979 colorado research in linguistics 5-1979 a study of fronting, voicing, and stopping based on olmsted's data in eighty-seven children mary coberly recommended citation tmp.1538088670.pdf.38qqs language origin and the nature of language: a linguistic interpretation of dante's de vulgari eloquentia 1 stong-jensen: language origin and the nature of language published by cu scholar, 1976 2 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 3 stong-jensen: language origin and the nature of language published by cu scholar, 1976 4 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 5 stong-jensen: language origin and the nature of language published by cu scholar, 1976 6 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 7 stong-jensen: language origin and the nature of language published by cu scholar, 1976 8 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 9 stong-jensen: language origin and the nature of language published by cu scholar, 1976 10 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 11 stong-jensen: language origin and the nature of language published by cu scholar, 1976 12 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 13 stong-jensen: language origin and the nature of language published by cu scholar, 1976 14 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 15 stong-jensen: language origin and the nature of language published by cu scholar, 1976 16 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/4 17 stong-jensen: language origin and the nature of language published by cu scholar, 1976 colorado research in linguistics 5-1976 language origin and the nature of language: a linguistic interpretation of dante's de vulgari eloquentia margaret stong-jensen recommended citation tmp.1538171221.pdf.qdz3o the ada verb of being in bahasa malaysia 1 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 2 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 3 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 4 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 5 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 6 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 7 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 8 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 9 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 10 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 11 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 12 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 13 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 14 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 15 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 16 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 17 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 18 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 19 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 20 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 21 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 22 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 23 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 24 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 25 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 26 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 27 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 28 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 29 mader: the ada verb of being in bahasa malaysia published by cu scholar, 1974 30 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/3 doi: https://doi.org/10.25810/eq7t-ct39 colorado research in linguistics 5-1974 the ada verb of being in bahasa malaysia robin mader recommended citation tmp.1538172200.pdf.1zj0y from lambert to zawadowski: a chronological microreview of excerpts on the relationship between linguistics and semiotics 1 romeo: from lambert to zawadowski published by cu scholar, 1977 2 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/4 3 romeo: from lambert to zawadowski published by cu scholar, 1977 4 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/4 colorado research in linguistics 5-1977 from lambert to zawadowski: a chronological microreview of excerpts on the relationship between linguistics and semiotics luigi romeo recommended citation tmp.1538168795.pdf._c8jt communicative function of intonation contour in proto-language: a short summary note 1 m en n: c om m un ic at iv e fu nc tio n of in to na tio n c on to ur in p ro to -l an gu ag e pu bl ish ed b y c u s ch ol ar , 1 98 9 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 0 [1 98 9] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 10 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ rd 9q -4 b3 7 3 m en n: c om m un ic at iv e fu nc tio n of in to na tio n c on to ur in p ro to -l an gu ag e pu bl ish ed b y c u s ch ol ar , 1 98 9 colorado research in linguistics 5-1989 communicative function of intonation contour in proto-language: a short summary note lise menn recommended citation tmp.1538225779.pdf.ig6gp against the distributional syllable 1 bell: against the distributional syllable published by cu scholar, 1972 2 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 3 bell: against the distributional syllable published by cu scholar, 1972 4 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 5 bell: against the distributional syllable published by cu scholar, 1972 6 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 7 bell: against the distributional syllable published by cu scholar, 1972 8 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 9 bell: against the distributional syllable published by cu scholar, 1972 10 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 11 bell: against the distributional syllable published by cu scholar, 1972 12 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 13 bell: against the distributional syllable published by cu scholar, 1972 14 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 15 bell: against the distributional syllable published by cu scholar, 1972 16 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 17 bell: against the distributional syllable published by cu scholar, 1972 18 colorado research in linguistics, vol. 2 [1972] https://scholar.colorado.edu/cril/vol2/iss1/1 doi: https://doi.org/10.25810/shkj-sv33 19 bell: against the distributional syllable published by cu scholar, 1972 colorado research in linguistics 10-1972 against the distributional syllable alan bell recommended citation tmp.1538173645.pdf.g6dxw some problems in the case grammar of awutu 1 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 2 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 3 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 4 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 5 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 6 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 7 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 8 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 9 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 10 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 11 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 12 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 13 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 14 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 15 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 16 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 17 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 18 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 19 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 20 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 21 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 22 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 23 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 24 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 25 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 26 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 27 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 28 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 29 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 30 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 31 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 32 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 33 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 34 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 35 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 36 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 37 frajzyngier: some problems in the case grammar of awutu published by cu scholar, 1973 38 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/2 doi: https://doi.org/10.25810/wwr4-yc16 colorado research in linguistics 5-1973 some problems in the case grammar of awutu zygmunt frajzyngier recommended citation tmp.1538173016.pdf.lorb0 the colorado university system for writing the lakhota language 1 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 2 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 3 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 4 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 5 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 6 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 7 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 8 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 9 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 10 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 11 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 12 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 13 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 14 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 15 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 16 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 17 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 18 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 19 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 20 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 21 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 22 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 23 taylor: the colorado university system for writing the lakhota language published by cu scholar, 1974 24 colorado research in linguistics, vol. 4 [1974] https://scholar.colorado.edu/cril/vol4/iss1/5 doi: https://doi.org/10.25810/6nnr-4f65 colorado research in linguistics 5-1974 the colorado university system for writing the lakhota language allan r. taylor recommended citation tmp.1538172431.pdf.y8a7i relative clauses in sesotho 1 khoali: relative clauses in sesotho published by cu scholar, 1979 2 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 3 khoali: relative clauses in sesotho published by cu scholar, 1979 4 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 5 khoali: relative clauses in sesotho published by cu scholar, 1979 6 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 7 khoali: relative clauses in sesotho published by cu scholar, 1979 8 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 9 khoali: relative clauses in sesotho published by cu scholar, 1979 10 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 11 khoali: relative clauses in sesotho published by cu scholar, 1979 12 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 13 khoali: relative clauses in sesotho published by cu scholar, 1979 14 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 15 khoali: relative clauses in sesotho published by cu scholar, 1979 16 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 17 khoali: relative clauses in sesotho published by cu scholar, 1979 18 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 19 khoali: relative clauses in sesotho published by cu scholar, 1979 20 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 21 khoali: relative clauses in sesotho published by cu scholar, 1979 22 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 23 khoali: relative clauses in sesotho published by cu scholar, 1979 24 colorado research in linguistics, vol. 8 [1979] https://scholar.colorado.edu/cril/vol8/iss1/3 doi: https://doi.org/10.25810/epaz-3e18 colorado research in linguistics 5-1979 relative clauses in sesotho ben t. khoali recommended citation tmp.1538088436.pdf.qqs7k synchronic applications for diachronic syntax: the grammaticalization of to be about to in english 1 jirsa: synchronic applications for diachronic syntax published by cu scholar, 1997 2 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/3 doi: https://doi.org/10.25810/5h5t-xg32 3 jirsa: synchronic applications for diachronic syntax published by cu scholar, 1997 4 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/3 doi: https://doi.org/10.25810/5h5t-xg32 5 jirsa: synchronic applications for diachronic syntax published by cu scholar, 1997 6 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/3 doi: https://doi.org/10.25810/5h5t-xg32 7 jirsa: synchronic applications for diachronic syntax published by cu scholar, 1997 colorado research in linguistics 1997 synchronic applications for diachronic syntax: the grammaticalization of to be about to in english bill jirsa recommended citation tmp.1537638431.pdf.sdol_ linguistics and mathematics: mix with care 1 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 2 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 3 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 4 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 5 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 6 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 7 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 8 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 9 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 10 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 11 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 12 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 13 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 14 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 15 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 16 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 17 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 18 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 19 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 20 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 21 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 22 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/1 23 abernathy: linguistics and mathematics: mix with care published by cu scholar, 1977 colorado research in linguistics 5-1977 linguistics and mathematics: mix with care robert abernathy recommended citation tmp.1538168437.pdf.ncnga gender and number agreement in standard arabic 1 a w ad : g en de r a nd n um be r a gr ee m en t i n st an da rd a ra bi c pu bl ish ed b y c u s ch ol ar , 1 99 0 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8b 2e -g 42 3 3 a w ad : g en de r a nd n um be r a gr ee m en t i n st an da rd a ra bi c pu bl ish ed b y c u s ch ol ar , 1 99 0 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8b 2e -g 42 3 colorado research in linguistics 5-1990 gender and number agreement in standard arabic maher awad recommended citation tmp.1538228725.pdf.nmwmx incorporation and kikamba verbal extensions 1 br ow n: in co rp or at io n an d k ik am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 4j 9w -9 51 0 3 br ow n: in co rp or at io n an d k ik am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 4j 9w -9 51 0 5 br ow n: in co rp or at io n an d k ik am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 4j 9w -9 51 0 7 br ow n: in co rp or at io n an d k ik am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 4j 9w -9 51 0 9 br ow n: in co rp or at io n an d k ik am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 10 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 4j 9w -9 51 0 colorado research in linguistics 5-1993 incorporation and kikamba verbal extensions linh-chan brown recommended citation tmp.1538231944.pdf.fjnbd conjunction, relativization, and complementation in persian 1 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 2 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 3 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 4 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 5 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 6 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 7 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 8 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 9 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 10 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 11 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 12 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 13 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 14 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 15 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 16 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 17 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 18 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 19 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 20 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 21 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 22 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 23 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 24 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 25 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 26 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 27 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 28 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 29 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 30 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 31 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 32 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 33 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 34 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 35 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 36 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 37 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 38 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 39 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 40 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 41 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 42 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 43 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 44 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 45 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 46 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 47 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 48 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 49 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 50 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 51 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 52 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 53 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 54 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 55 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 56 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 57 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 58 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 59 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 60 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 61 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 62 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 63 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 64 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 65 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 66 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 67 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 68 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 69 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 70 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 71 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 72 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 73 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 74 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 75 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 76 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 77 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 78 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 79 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 80 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 81 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 82 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 83 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 84 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 85 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 86 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 87 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 88 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 89 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 90 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 91 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 92 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 93 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 94 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 95 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 96 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 97 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 98 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 99 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 100 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 101 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 102 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 103 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 104 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 105 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 106 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 107 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 108 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 109 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 110 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 111 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 112 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 113 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 114 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 115 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 116 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 117 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 118 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 119 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 120 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 121 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 122 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 123 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 124 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 125 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 126 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 127 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 128 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 129 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 130 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 131 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 132 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 133 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 134 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 135 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 136 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 137 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 138 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 139 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 140 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 141 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 142 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 143 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 144 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 145 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 146 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 147 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 148 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 149 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 150 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 151 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 152 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 153 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 154 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 155 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 156 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 157 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 158 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 159 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 160 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 161 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 162 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 163 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 164 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 165 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 166 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 167 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 168 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 169 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 170 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 171 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 172 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 173 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 174 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 175 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 176 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 177 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 178 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 179 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 180 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 181 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 182 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 183 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 184 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 185 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 186 colorado research in linguistics, vol. 5 [1975] https://scholar.colorado.edu/cril/vol5/iss1/2 187 tabaian: conjunction, relativization, and complementation in persian published by cu scholar, 1975 colorado research in linguistics 5-1975 conjunction, relativization, and complementation in persian hessam tabaian recommended citation tmp.1538170345.pdf.xx_qg onset-rhyme temporal structure of mandarin syllables 1 be ll an d li u: o ns et -r hy m e te m po ra l s tr uc tu re o f m an da rin s yl la bl es pu bl ish ed b y c u s ch ol ar , 1 99 0 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ v1 gq -t p5 6 3 be ll an d li u: o ns et -r hy m e te m po ra l s tr uc tu re o f m an da rin s yl la bl es pu bl ish ed b y c u s ch ol ar , 1 99 0 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ v1 gq -t p5 6 5 be ll an d li u: o ns et -r hy m e te m po ra l s tr uc tu re o f m an da rin s yl la bl es pu bl ish ed b y c u s ch ol ar , 1 99 0 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ v1 gq -t p5 6 7 be ll an d li u: o ns et -r hy m e te m po ra l s tr uc tu re o f m an da rin s yl la bl es pu bl ish ed b y c u s ch ol ar , 1 99 0 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ v1 gq -t p5 6 9 be ll an d li u: o ns et -r hy m e te m po ra l s tr uc tu re o f m an da rin s yl la bl es pu bl ish ed b y c u s ch ol ar , 1 99 0 colorado research in linguistics 5-1990 onset-rhyme temporal structure of mandarin syllables alan bell meichun liu recommended citation tmp.1538228828.pdf.bwmqc morphological coding and syntactic role in the grammar of panjabi complex sentences 1 c lif fo rd : m or ph ol og ic al c od in g an d sy nt ac tic r ol e in th e g ra m m ar o f p an ja bi c om pl ex s en te nc es pu bl ish ed b y c u s ch ol ar , 1 99 0 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ qz bg -p n6 4 3 c lif fo rd : m or ph ol og ic al c od in g an d sy nt ac tic r ol e in th e g ra m m ar o f p an ja bi c om pl ex s en te nc es pu bl ish ed b y c u s ch ol ar , 1 99 0 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ qz bg -p n6 4 5 c lif fo rd : m or ph ol og ic al c od in g an d sy nt ac tic r ol e in th e g ra m m ar o f p an ja bi c om pl ex s en te nc es pu bl ish ed b y c u s ch ol ar , 1 99 0 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ qz bg -p n6 4 7 c lif fo rd : m or ph ol og ic al c od in g an d sy nt ac tic r ol e in th e g ra m m ar o f p an ja bi c om pl ex s en te nc es pu bl ish ed b y c u s ch ol ar , 1 99 0 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ qz bg -p n6 4 9 c lif fo rd : m or ph ol og ic al c od in g an d sy nt ac tic r ol e in th e g ra m m ar o f p an ja bi c om pl ex s en te nc es pu bl ish ed b y c u s ch ol ar , 1 99 0 10 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ qz bg -p n6 4 colorado research in linguistics 5-1990 morphological coding and syntactic role in the grammar of panjabi complex sentences joseph clifford recommended citation tmp.1538228926.pdf.s4jgm balto-slavic or baltic and slavic 1 fi sh er : b al to -s la vi c or b al tic a nd s la vi c pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ az qe -y v7 5 3 fi sh er : b al to -s la vi c or b al tic a nd s la vi c pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ az qe -y v7 5 5 fi sh er : b al to -s la vi c or b al tic a nd s la vi c pu bl ish ed b y c u s ch ol ar , 1 99 3 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ az qe -y v7 5 7 fi sh er : b al to -s la vi c or b al tic a nd s la vi c pu bl ish ed b y c u s ch ol ar , 1 99 3 colorado research in linguistics 5-1993 balto-slavic or baltic and slavic julia fisher recommended citation tmp.1538231679.pdf.zkznu verbal affixation and grammatical relations in modern standard indonesian 1 sh ay : v er ba l a ffi xa tio n an d g ra m m at ic al r el at io ns in m od er n st an da rd in do ne sia n pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /8 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ df ap -p q3 6 3 sh ay : v er ba l a ffi xa tio n an d g ra m m at ic al r el at io ns in m od er n st an da rd in do ne sia n pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /8 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ df ap -p q3 6 5 sh ay : v er ba l a ffi xa tio n an d g ra m m at ic al r el at io ns in m od er n st an da rd in do ne sia n pu bl ish ed b y c u s ch ol ar , 1 99 3 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /8 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ df ap -p q3 6 7 sh ay : v er ba l a ffi xa tio n an d g ra m m at ic al r el at io ns in m od er n st an da rd in do ne sia n pu bl ish ed b y c u s ch ol ar , 1 99 3 colorado research in linguistics 5-1993 verbal affixation and grammatical relations in modern standard indonesian erin shay recommended citation tmp.1538232024.pdf.3pxx5 naming things for children: the basic level is not ad hoc 1 wallace ross: naming things for children published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/8 doi: https://doi.org/10.25810/3wt2-zr69 3 wallace ross: naming things for children published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/8 doi: https://doi.org/10.25810/3wt2-zr69 5 wallace ross: naming things for children published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/8 doi: https://doi.org/10.25810/3wt2-zr69 7 wallace ross: naming things for children published by cu scholar, 1995 8 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/8 doi: https://doi.org/10.25810/3wt2-zr69 9 wallace ross: naming things for children published by cu scholar, 1995 10 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/8 doi: https://doi.org/10.25810/3wt2-zr69 11 wallace ross: naming things for children published by cu scholar, 1995 12 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/8 doi: https://doi.org/10.25810/3wt2-zr69 colorado research in linguistics 1995 naming things for children: the basic level is not ad hoc valerie wallace ross recommended citation tmp.1537648672.pdf.ah9_c reference-tracking system and anaphora in mandarin chinese conversational discourse 1 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 74 ak -3 v7 1 3 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 74 ak -3 v7 1 5 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 74 ak -3 v7 1 7 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 74 ak -3 v7 1 9 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 10 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 74 ak -3 v7 1 11 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 12 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 74 ak -3 v7 1 13 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 14 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 74 ak -3 v7 1 15 ta o: r ef er en ce -t ra ck in g sy st em a nd a na ph or a in m an da rin c hi ne se c on ve rs at io na l d isc ou rs e pu bl ish ed b y c u s ch ol ar , 1 99 0 colorado research in linguistics 5-1990 reference-tracking system and anaphora in mandarin chinese conversational discourse liang tao recommended citation tmp.1538229324.pdf.yj2sv second person deixis in japanese and power semantics 1 ta ka ha ra : s ec on d pe rs on d ei xi s i n ja pa ne se a nd p ow er s em an tic s pu bl ish ed b y c u s ch ol ar , 1 99 0 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 1a fq -t d0 2 3 ta ka ha ra : s ec on d pe rs on d ei xi s i n ja pa ne se a nd p ow er s em an tic s pu bl ish ed b y c u s ch ol ar , 1 99 0 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 1a fq -t d0 2 5 ta ka ha ra : s ec on d pe rs on d ei xi s i n ja pa ne se a nd p ow er s em an tic s pu bl ish ed b y c u s ch ol ar , 1 99 0 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 1a fq -t d0 2 colorado research in linguistics 5-1990 second person deixis in japanese and power semantics kumiko takahara recommended citation tmp.1538229230.pdf.edny9 a typological investigation of the structure of consonant inventories 1 ra ym on d: a t yp ol og ic al in ve st ig at io n of th e st ru ct ur e of c on so na nt in ve nt or ie s pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8d 78 -9 s9 2 3 ra ym on d: a t yp ol og ic al in ve st ig at io n of th e st ru ct ur e of c on so na nt in ve nt or ie s pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8d 78 -9 s9 2 5 ra ym on d: a t yp ol og ic al in ve st ig at io n of th e st ru ct ur e of c on so na nt in ve nt or ie s pu bl ish ed b y c u s ch ol ar , 1 99 3 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8d 78 -9 s9 2 7 ra ym on d: a t yp ol og ic al in ve st ig at io n of th e st ru ct ur e of c on so na nt in ve nt or ie s pu bl ish ed b y c u s ch ol ar , 1 99 3 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8d 78 -9 s9 2 9 ra ym on d: a t yp ol og ic al in ve st ig at io n of th e st ru ct ur e of c on so na nt in ve nt or ie s pu bl ish ed b y c u s ch ol ar , 1 99 3 10 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8d 78 -9 s9 2 11 ra ym on d: a t yp ol og ic al in ve st ig at io n of th e st ru ct ur e of c on so na nt in ve nt or ie s pu bl ish ed b y c u s ch ol ar , 1 99 3 12 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /1 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 8d 78 -9 s9 2 13 ra ym on d: a t yp ol og ic al in ve st ig at io n of th e st ru ct ur e of c on so na nt in ve nt or ie s pu bl ish ed b y c u s ch ol ar , 1 99 3 colorado research in linguistics 5-1993 a typological investigation of the structure of consonant inventories william d. raymond recommended citation tmp.1538231438.pdf.qblgv wichita text structure 1 ro od : w ic hi ta t ex t s tr uc tu re pu bl ish ed b y c u s ch ol ar , 1 98 6 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ m ke cfj0 7 3 ro od : w ic hi ta t ex t s tr uc tu re pu bl ish ed b y c u s ch ol ar , 1 98 6 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 7 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ m ke cfj0 7 5 ro od : w ic hi ta t ex t s tr uc tu re pu bl ish ed b y c u s ch ol ar , 1 98 6 colorado research in linguistics 1986 wichita text structure david s. rood recommended citation tmp.1538224387.pdf.gwnla what was funny? discourse referents in pronoun use of young children 1 bu rn s: w ha t w as f un ny ? d isc ou rs e re fe re nt s i n pr on ou n u se o f y ou ng c hi ld re n pu bl ish ed b y c u s ch ol ar , 1 98 6 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 4v t8 -x x7 4 3 bu rn s: w ha t w as f un ny ? d isc ou rs e re fe re nt s i n pr on ou n u se o f y ou ng c hi ld re n pu bl ish ed b y c u s ch ol ar , 1 98 6 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 9 [1 98 6] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 9/ iss 1/ 3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 4v t8 -x x7 4 5 bu rn s: w ha t w as f un ny ? d isc ou rs e re fe re nt s i n pr on ou n u se o f y ou ng c hi ld re n pu bl ish ed b y c u s ch ol ar , 1 98 6 colorado research in linguistics 1986 what was funny? discourse referents in pronoun use of young children rebecca burns recommended citation tmp.1538224023.pdf.9twu6 repair strategies in conversational kickapoo 1 gomez de garcia: repair strategies in conversational kickapoo published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/4 doi: https://doi.org/10.25810/x81d-mc97 3 gomez de garcia: repair strategies in conversational kickapoo published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/4 doi: https://doi.org/10.25810/x81d-mc97 5 gomez de garcia: repair strategies in conversational kickapoo published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/4 doi: https://doi.org/10.25810/x81d-mc97 7 gomez de garcia: repair strategies in conversational kickapoo published by cu scholar, 1995 8 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/4 doi: https://doi.org/10.25810/x81d-mc97 9 gomez de garcia: repair strategies in conversational kickapoo published by cu scholar, 1995 10 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/4 doi: https://doi.org/10.25810/x81d-mc97 colorado research in linguistics 1995 repair strategies in conversational kickapoo jule gomez de garcia recommended citation tmp.1537644616.pdf.bcbp2 grammaticalization of aleyn in yiddish 1 halperin biasca: grammaticalization of aleyn in yiddish published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/3 doi: https://doi.org/10.25810/a2f6-wy92 3 halperin biasca: grammaticalization of aleyn in yiddish published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/3 doi: https://doi.org/10.25810/a2f6-wy92 5 halperin biasca: grammaticalization of aleyn in yiddish published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/3 doi: https://doi.org/10.25810/a2f6-wy92 7 halperin biasca: grammaticalization of aleyn in yiddish published by cu scholar, 1995 8 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/3 doi: https://doi.org/10.25810/a2f6-wy92 9 halperin biasca: grammaticalization of aleyn in yiddish published by cu scholar, 1995 10 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/3 doi: https://doi.org/10.25810/a2f6-wy92 11 halperin biasca: grammaticalization of aleyn in yiddish published by cu scholar, 1995 12 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/3 doi: https://doi.org/10.25810/a2f6-wy92 13 halperin biasca: grammaticalization of aleyn in yiddish published by cu scholar, 1995 14 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/3 doi: https://doi.org/10.25810/a2f6-wy92 colorado research in linguistics 1995 grammaticalization of aleyn in yiddish debra halperin biasca recommended citation tmp.1537644941.pdf.ibspv the scope of intransitivity in basque 1 bellver: the scope of intransitivity in basque published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 3 bellver: the scope of intransitivity in basque published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 5 bellver: the scope of intransitivity in basque published by cu scholar, 1994 6 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 7 bellver: the scope of intransitivity in basque published by cu scholar, 1994 8 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 9 bellver: the scope of intransitivity in basque published by cu scholar, 1994 10 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 11 bellver: the scope of intransitivity in basque published by cu scholar, 1994 12 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 13 bellver: the scope of intransitivity in basque published by cu scholar, 1994 14 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 15 bellver: the scope of intransitivity in basque published by cu scholar, 1994 16 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 17 bellver: the scope of intransitivity in basque published by cu scholar, 1994 18 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 19 bellver: the scope of intransitivity in basque published by cu scholar, 1994 20 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/3 doi: https://doi.org/10.25810/bqsj-6r58 21 bellver: the scope of intransitivity in basque published by cu scholar, 1994 colorado research in linguistics 1994 the scope of intransitivity in basque phyllis bellver recommended citation tmp.1537747646.pdf.dyt2d siouan linguistics: an assessment of where we are 1 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 2 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 3 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 4 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 5 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 6 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 7 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 8 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 9 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 10 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 11 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 12 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 13 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 14 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 15 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 16 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 17 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 18 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 19 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 20 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 21 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 22 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 23 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 24 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 25 rood: siouan linguistics: an assessment of where we are published by cu scholar, 1977 26 colorado research in linguistics, vol. 7 [1977] https://scholar.colorado.edu/cril/vol7/iss1/5 colorado research in linguistics 5-1977 siouan linguistics: an assessment of where we are david rood recommended citation tmp.1538168947.pdf.mlq1f peter ramus (1515-1572) as the first 'modern' structuralist 1 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 2 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 3 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 4 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 5 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 6 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 7 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 8 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 9 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 10 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 11 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 12 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 13 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 14 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 15 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 16 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 17 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 18 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 19 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 20 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 21 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 22 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 23 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 24 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 25 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 26 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 27 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 28 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 29 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 30 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 31 breva-claramonte: peter ramus (1515-1572) as the first 'modern' structuralist published by cu scholar, 1976 32 colorado research in linguistics, vol. 6 [1976] https://scholar.colorado.edu/cril/vol6/iss1/1 colorado research in linguistics 5-1976 peter ramus (1515-1572) as the first 'modern' structuralist manuel breva-claramonte recommended citation tmp.1538170659.pdf.myp44 provencal sor and molher, lone survivors of the feminine imparisyllables 1 jensen: provencal sor and molher, lone survivors of the feminine imparisyllables published by cu scholar, 1973 2 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/3 doi: https://doi.org/10.25810/1z38-ex85 3 jensen: provencal sor and molher, lone survivors of the feminine imparisyllables published by cu scholar, 1973 4 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/3 doi: https://doi.org/10.25810/1z38-ex85 5 jensen: provencal sor and molher, lone survivors of the feminine imparisyllables published by cu scholar, 1973 6 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/3 doi: https://doi.org/10.25810/1z38-ex85 7 jensen: provencal sor and molher, lone survivors of the feminine imparisyllables published by cu scholar, 1973 8 colorado research in linguistics, vol. 3 [1973] https://scholar.colorado.edu/cril/vol3/iss1/3 doi: https://doi.org/10.25810/1z38-ex85 colorado research in linguistics 5-1973 provencal sor and molher, lone survivors of the feminine imparisyllables frede jensen recommended citation tmp.1538173123.pdf.pqftc letter detection in german silent reading: some linguistic issues 1 bu ck -g en gl er : l et te r d et ec tio n in g er m an s ile nt r ea di ng : s om e li ng ui st ic is su es pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 65 19 -y t4 7 3 bu ck -g en gl er : l et te r d et ec tio n in g er m an s ile nt r ea di ng : s om e li ng ui st ic is su es pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 65 19 -y t4 7 5 bu ck -g en gl er : l et te r d et ec tio n in g er m an s ile nt r ea di ng : s om e li ng ui st ic is su es pu bl ish ed b y c u s ch ol ar , 1 99 3 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 65 19 -y t4 7 7 bu ck -g en gl er : l et te r d et ec tio n in g er m an s ile nt r ea di ng : s om e li ng ui st ic is su es pu bl ish ed b y c u s ch ol ar , 1 99 3 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /3 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 65 19 -y t4 7 9 bu ck -g en gl er : l et te r d et ec tio n in g er m an s ile nt r ea di ng : s om e li ng ui st ic is su es pu bl ish ed b y c u s ch ol ar , 1 99 3 colorado research in linguistics 5-1993 letter detection in german silent reading: some linguistic issues carolyn buck-gengler recommended citation tmp.1538231592.pdf.mxbg0 the preferred argument structure in mandarin chinese and its cognitive implications 1 lo ng : t he p re fe rr ed a rg um en t s tr uc tu re in m an da rin c hi ne se a nd it s c og ni tiv e im pl ic at io ns pu bl ish ed b y c u s ch ol ar , 1 99 0 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ k4 bt -4 q4 6 3 lo ng : t he p re fe rr ed a rg um en t s tr uc tu re in m an da rin c hi ne se a nd it s c og ni tiv e im pl ic at io ns pu bl ish ed b y c u s ch ol ar , 1 99 0 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ k4 bt -4 q4 6 5 lo ng : t he p re fe rr ed a rg um en t s tr uc tu re in m an da rin c hi ne se a nd it s c og ni tiv e im pl ic at io ns pu bl ish ed b y c u s ch ol ar , 1 99 0 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ k4 bt -4 q4 6 7 lo ng : t he p re fe rr ed a rg um en t s tr uc tu re in m an da rin c hi ne se a nd it s c og ni tiv e im pl ic at io ns pu bl ish ed b y c u s ch ol ar , 1 99 0 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /4 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ k4 bt -4 q4 6 9 lo ng : t he p re fe rr ed a rg um en t s tr uc tu re in m an da rin c hi ne se a nd it s c og ni tiv e im pl ic at io ns pu bl ish ed b y c u s ch ol ar , 1 99 0 colorado research in linguistics 5-1990 the preferred argument structure in mandarin chinese and its cognitive implications zhihua long recommended citation tmp.1538229036.pdf.mxvxq complementation and modality: two complementizers in east dangla 1 shay: complementation and modality published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/7 doi: https://doi.org/10.25810/cjw2-r670 3 shay: complementation and modality published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/7 doi: https://doi.org/10.25810/cjw2-r670 5 shay: complementation and modality published by cu scholar, 1994 6 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/7 doi: https://doi.org/10.25810/cjw2-r670 7 shay: complementation and modality published by cu scholar, 1994 8 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/7 doi: https://doi.org/10.25810/cjw2-r670 9 shay: complementation and modality published by cu scholar, 1994 10 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/7 doi: https://doi.org/10.25810/cjw2-r670 11 shay: complementation and modality published by cu scholar, 1994 12 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/7 doi: https://doi.org/10.25810/cjw2-r670 13 shay: complementation and modality published by cu scholar, 1994 colorado research in linguistics 1994 complementation and modality: two complementizers in east dangla erin shay recommended citation tmp.1537749846.pdf.htclc the syntax and semantics of complement clauses in arabic 1 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 2 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 3 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 4 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 5 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 6 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 7 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 8 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 9 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 10 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 11 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 12 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 13 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 14 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 15 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 16 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 17 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 18 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 19 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 20 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 21 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 22 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 23 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 24 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 25 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 26 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 27 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 28 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/1 doi: https://doi.org/10.25810/6fv3-q762 29 awad: the syntax and semantics of complement clauses in arabic published by cu scholar, 1998 colorado research in linguistics 1998 the syntax and semantics of complement clauses in arabic maher awad recommended citation tmp.1537465929.pdf.wquj8 applying optimality theory to german phonology: [x]/[ç] distribution and final devoicing 1 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/4 3 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/4 5 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 6 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/4 7 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 8 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/4 9 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 10 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/4 11 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 12 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/4 13 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 14 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/4 15 buck-gengler: applying optimality theory to german phonology published by cu scholar, 1994 colorado research in linguistics 1994 applying optimality theory to german phonology: [x]/[ç] distribution and final devoicing carolyn buck-gengler recommended citation tmp.1537748207.pdf.qnmnf mora-based temporal adjustments in japanese 1 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/1 doi: https://doi.org/10.25810/2ddh-9161 3 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/1 doi: https://doi.org/10.25810/2ddh-9161 5 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 6 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/1 doi: https://doi.org/10.25810/2ddh-9161 7 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 8 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/1 doi: https://doi.org/10.25810/2ddh-9161 9 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 10 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/1 doi: https://doi.org/10.25810/2ddh-9161 11 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 12 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/1 doi: https://doi.org/10.25810/2ddh-9161 13 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 14 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/1 doi: https://doi.org/10.25810/2ddh-9161 15 asano: mora-based temporal adjustments in japanese published by cu scholar, 1994 colorado research in linguistics 1994 mora-based temporal adjustments in japanese yoshiteru asano recommended citation tmp.1537746990.pdf.hmsct english stream names and linguistic stratification: a test of nicolaisen's geographic model 1 n ol te : e ng lis h st re am n am es a nd l in gu ist ic s tr at ifi ca tio n: a t es t o f n ic ol ai se n' s g eo gr ap hi c m od el pu bl ish ed b y c u s ch ol ar , 1 99 0 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ ne qe -h k7 0 3 n ol te : e ng lis h st re am n am es a nd l in gu ist ic s tr at ifi ca tio n: a t es t o f n ic ol ai se n' s g eo gr ap hi c m od el pu bl ish ed b y c u s ch ol ar , 1 99 0 4 colorado research in linguistics, v ol. 11 [1990] https://scholar.colorado.edu/cril/vol11/iss1/5 d o i: https://doi.org/10.25810/neqe-hk70 5 n ol te : e ng lis h st re am n am es a nd l in gu ist ic s tr at ifi ca tio n: a t es t o f n ic ol ai se n' s g eo gr ap hi c m od el pu bl ish ed b y c u s ch ol ar , 1 99 0 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 1 [1 99 0] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 11 /is s1 /5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ ne qe -h k7 0 colorado research in linguistics 5-1990 english stream names and linguistic stratification: a test of nicolaisen's geographic model nancy nolte recommended citation tmp.1538229157.pdf.99wvi argument gapping in coordinate conjunction constructions in japanese 1 asano: argument gapping in coordinate conjunction constructions in japanese published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/1 doi: https://doi.org/10.25810/gr5s-e264 3 asano: argument gapping in coordinate conjunction constructions in japanese published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/1 doi: https://doi.org/10.25810/gr5s-e264 5 asano: argument gapping in coordinate conjunction constructions in japanese published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/1 doi: https://doi.org/10.25810/gr5s-e264 7 asano: argument gapping in coordinate conjunction constructions in japanese published by cu scholar, 1995 8 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/1 doi: https://doi.org/10.25810/gr5s-e264 9 asano: argument gapping in coordinate conjunction constructions in japanese published by cu scholar, 1995 colorado research in linguistics 1995 argument gapping in coordinate conjunction constructions in japanese yoshiteru asano recommended citation tmp.1537643564.pdf.ezzs6 what does that refer to? non-personal pronoun anaphora in english conversational discourse 1 sparks: what does that refer to? published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/8 doi: https://doi.org/10.25810/sqsy-fb89 3 sparks: what does that refer to? published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/8 doi: https://doi.org/10.25810/sqsy-fb89 5 sparks: what does that refer to? published by cu scholar, 1994 6 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/8 doi: https://doi.org/10.25810/sqsy-fb89 7 sparks: what does that refer to? published by cu scholar, 1994 8 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/8 doi: https://doi.org/10.25810/sqsy-fb89 9 sparks: what does that refer to? published by cu scholar, 1994 10 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/8 doi: https://doi.org/10.25810/sqsy-fb89 11 sparks: what does that refer to? published by cu scholar, 1994 12 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/8 doi: https://doi.org/10.25810/sqsy-fb89 colorado research in linguistics 1994 what does that refer to? non-personal pronoun anaphora in english conversational discourse randall b. sparks recommended citation tmp.1537750339.pdf.a1nkv acquisition of english interjections ouch, yuck, and oops in early childhood 1 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 2 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/1 doi: https://doi.org/10.25810/6t1p-7625 3 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 4 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/1 doi: https://doi.org/10.25810/6t1p-7625 5 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 6 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/1 doi: https://doi.org/10.25810/6t1p-7625 7 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 8 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/1 doi: https://doi.org/10.25810/6t1p-7625 9 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 10 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/1 doi: https://doi.org/10.25810/6t1p-7625 11 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 12 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/1 doi: https://doi.org/10.25810/6t1p-7625 13 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 14 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/1 doi: https://doi.org/10.25810/6t1p-7625 15 asano: acquisition of english interjections ouch, yuck, and oops in early childhood published by cu scholar, 1997 colorado research in linguistics 1997 acquisition of english interjections ouch, yuck, and oops in early childhood yoshiteru asano recommended citation tmp.1537634975.pdf.kemn6 the functions of kamba verbal extensions 1 a ns ch ut z: th e fu nc tio ns o f k am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 85 0z -9 a5 7 3 a ns ch ut z: th e fu nc tio ns o f k am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 85 0z -9 a5 7 5 a ns ch ut z: th e fu nc tio ns o f k am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 6 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 85 0z -9 a5 7 7 a ns ch ut z: th e fu nc tio ns o f k am ba v er ba l e xt en sio ns pu bl ish ed b y c u s ch ol ar , 1 99 3 8 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /6 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ 85 0z -9 a5 7 colorado research in linguistics 5-1993 the functions of kamba verbal extensions arlea anschutz recommended citation tmp.1538231855.pdf.0hdjv language identification and language specific letter-to-sound rules language identification and language specific letter-to-sound rules stephen lewis, katie mcgrath, jeffrey reuppel university of colorado at boulder this paper describes a system that improves automatic arpabet transcription by addressing performance issues resulting from arabic and russian transliteration in english text. our system is called ear (english, arabic, russian). the ear system has two components: 1. an n-gram language identifier module which classifies an incoming unknown word as arabic, russian, or english, 2. language specific letter to sound rules which output a pronunciation for a word based on its classification. our results show overall system error reduction rates at upwards of 45% as compared to a system trained only on english. 1. introduction the sparsity of transcribed conversational english makes assembling a corpus for speech recognition training a challenging task. one alternative resource for dealing with sparsity is to mine the world wide web for transcribed conversations. utilizing this resource though poses a number of new problems from text normalization to producing optimized output for letter to speech. this transcribed english text, especially news text, contains significant a number of transliterated foreign proper nouns. without normalization the output of a text-to-speech system will often generate inappropriate pronunciations because the letter-to-sound rules have been trained on english spelling standards, not on foreign transliteration standards. applying english letter-to-sound rules to arabic transliteration, for instance, can result in the following mispronunciation. example 1. hibaaq hh ay b ae kd1 to address the issue of poorly produced pronunciation data due to the presence of non-english words, we have implemented a two-tiered language classifier. the first tier of this classifier serves to identify non-english words in the conversational data scraped from the web. the second tier processes the results of this classification through specific letter-to-sound rules that have been trained for each non-english language in question. more specifically, our research has focused on solving these problems for arabic and russian words in our web scraped corpus. words from both of these languages had 1 wrong letter to sound output of arabic word expanded into arbabet symbols using decision trees trained on cmu’s english lexicon. for details about arpabet symbols, see (jurafsky & martin 2000, p 94-95). arpabet [hh ay b ae kd] is roughly equivalent to ipa [����� �]. colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by lewis, stephen; mcgrath, katie; reuppel, jeffrey. 1 lewis et al.: language identification and language specific letter-to-sound rules published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 already been identified as both common and problematic to the production of training material. 2. past/related work off the shelf technology which can perform language identification on text in general is available. these same tools can also be used to produce letter to speech pronunciation output for non-english words. however none of these systems are appropriate for the task that we wish to accomplish. most existing language classification systems use the simple trick of determining the character encoding of a document to perform language identification. a few systems use n-gram statistics at the word level to perform this task. both of these techniques can be used to identify non-english words successfully. however this identification works within the context of the entirety of the text in question being composed in the non-english language. this type of system is not optimized to deal with identifying the origin of non-english words within an otherwise english language text. the same is often true of letter-to-sound systems. they are built to pronounce nonenglish words correctly within the context of the language of origin. however the pronunciations are representative of native speaker pronunciation and inappropriate to producing the letter to sound output indicative of a monolingual english speaker attempting to pronounce an unknown word of foreign origin. 3. language classification 3.1. training data since arabic and russian proper nouns had previously been identified as the primary cause for pronunciation errors our first task was to acquire training data by which to identify these words. using the world wide web as our resource we constructed a database of transliterated proper nouns in arabic and russian names. the arabic names were obtained on the web in numerous locations to create a list of 3143 unique arabic names. the russian names were all obtained from a single web site that provided 20577 unique russian names (goldschmidt 1996). in order to create a standard of comparison for english words, we also collected a list of the 10,000 most common words longer than 4 letters from the brown corpus. we collected english words from the brown corpus with the belief that such an assemblage would better represent the body of unknown english words. while new foreign words are generally proper nouns, "unknown" english words tend to be morphological variants of words in the lexicon. in addition, english first names are not at all representative of the greater language, as evidenced by the 5 most common 4-grams from a list of english names and the brown corpus (table 1). 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/6 doi: https://doi.org/10.25810/60mf-fn94 http://www.sca.org/heraldry/paul/ language identification and language specific letter-to-sound rules 3 names brown corpus mar ing ana tion nna ion ina atio anna ted table 1. the 5 most common 4-grams from a list of english names and from the brown corpus 3.2. classification algorithm we implemented an n-gram classifier to handle language type identification. each training word was segmented into individual letters. individual 4-grams were constructed using each four-letter set. in addition much like sentence boundaries are marked, letters which begin and end words were marked with and <\s>. each individual 4-gram was then assigned a specific probability based on frequency. the word to be classified was also segmented into 4-grams and then labeled for language using the following equation (equation 1). equation 1: c = argmax c p(c|x1,…,xn) = argmax c p(c) π p(xi|c) i=1 c – final language classification c – individual language classification x – 4-gram n – number of 4-grams in the word being classified prior probabilities for each language were generated in proportion to the content of the cnn corpus. this set of news transcriptions was then compared against the cmu lexicon. this comparison returned a list of 1001 "unknown" words. each unknown word was labeled by a team of linguistics graduate students as being arabic, russian, or other. each word was then classified according to the majority label given it and the percentages of each language classification determined the priors. the priors were set as in table 2: 3 lewis et al.: language identification and language specific letter-to-sound rules published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 english: 0.805 arabic: 0.156 russian: 0.039 table 2. prior probabilities for labeling words english, arabic, or russian as with any n-gram classifier, we needed a way to calculate probabilities for new 4grams that never appeared in the training data. to account for these unseen 4-grams, we used a modified add-one smoothing method using this method, the probability for each unseen 4-gram was assigned to be the same as that of the 4-grams with the lowest overall probability. in the same 1001 unknown words mentioned above, unseen 4-grams accounted for only 0.126% of all 4-grams. as smoothing is invoked so infrequently, more sophisticated forms of smoothing or back-off do not seem to be fertile avenues to travel down for system improvement. 4. letter to sound rules 4.1. training data our letter to sound output systems for arabic and russian are built upon the integration of two separate resources. first our lexicon is structured in the same way as the cmu pronunciation dictionary reformatted to sphinx format (cmu sphinx). second we have used the sonic (pellom, 2003) decision tree software to train our the letter-tosound rules on producing the correct output consisting of a word and it's arpabet transcription. given the absence of a proper corpus of arabic and russian words transcribed in this manner, we once again scraped the web for data. after collecting transliterated words from various resources theses words were hand-transcribed into arpabet phonemic output by a team of graduate linguists. in all we built two corpora of transliterated words, 844 russian words and 582 arabic words. this corpus is small, but since the words were hand-transcribed, its data is at least reliable. both transcription time and the previous shortage of appropriately transliterated words contributed to the decision to use a small but reliable data set. 4.2. training algorithm for our decision tree training algorithm we used borrowed technology from the sonic system. the sonic system allows us to produce letter-to-sound output for words which we have never seen before and which do not exist in our language specific training corpora. 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/6 doi: https://doi.org/10.25810/60mf-fn94 http://fife.speech.cs.cmu.edu/sphinx/ language identification and language specific letter-to-sound rules 5 the algorithm works by extracting feature vectors from the input data consisting of the center letter plus 3 letters of context. each output phoneme is selected using a greatest reduction of entropy measure. data preparation for using the sonic system requires input and produced output as shown in example 2. example 2: input golova g ow l ax v aa output k aa r iy k uw (caricu) 5. results/analysis analysis of the system was done in 3 distinct phases – analysis of the classifier, analysis of the letter-to-sound rules, and analysis of the complete system 5.1. phase 1: analysis of the classifier the language identification classifier training data was split to create test data from 20% of each of the three language word lists. the classifier was then trained using the remaining 80% of each list. this process was repeated 4 times with the training data split differently each time, providing 5 different overlapping training sets and test sets. each was analyzed for precision by simply counting the number of times each classifier correctly labeled the words in the test sets for each language. the classifier's precision on the arabic test data sets described above ranged from .86 to 1.0 with a mean of .92. on the russian data precision ranged from .79 to .83 with a mean of .81. the classifier achieved its highest accuracy and greatest consistency on the english data with a precision ranging from .986 to .992 and a mean of .988. 5 lewis et al.: language identification and language specific letter-to-sound rules published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 language identification performance 0.5 0.55 0.6 0.65 0.7 0.75 0.8 0.85 0.9 0.95 1 classifier 1 0.998 0.988 0.804 classifier 2 0.87 0.987 0.833 classifier 3 1 0.992 0.829 classifier 4 0.874 0.989 0.801 classifier 5 0.855 0.988 0.79 mean 0.92 0.989 0.811 arabic english russian figure 1. performance of the classifier as trained and tested on 5 overlapping data sets (classifiers 1-5) 5.2. phase 2: analysis of the letter-to-sound rules of the hand transcribed arabic and russian transliterated words, 10% were reserved for testing the language specific letter-to-sound rule decision trees. performance of the language specific decision trees was compared against the performance of a decision tree trained on the cmu lexicon as a baseline. accuracy was determined by counting the phones that matched between the hand transcription and the decision tree transcription. the russian decision tree performed at a precision of .93 (error rate .07) on the russian test set. the baseline decision tree performed at a precision of .66 (error rate .33) on the 0.5 0.7 0.9 1.1 letter-to-sound performance baseline (cmu) 0.711 0.662 language specific 0.877 0.928 arabic russian figure 2. performance of language specific letter to sound rules on language specific data as compared to baseline 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/6 doi: https://doi.org/10.25810/60mf-fn94 language identification and language specific letter-to-sound rules 7 same test set, showing a 79 percent reduction in error rate by using the russian decision tree. the arabic decision tree performed at a precision of .88 (error rate .12) on the arabic test set. the baseline decision tree performed at a precision of .71 (error rate .29) on the same test set, showing a 57 percent reduction in error rate by using the arabic decision tree. 5.3. phase 3: analysis of the complete system to test the complete system, we isolated 50 words from each hand transcribed test set. we then isolated 50 english words from the cmu lexicon, and combined the three sets, making 150 transcribed words, 50 in each language. these words were classified using the language identification classifier and then transcribed using appropriate decision trees based on the classifier labels. counts were then made of the phones in the system's transcription that matched the hand transcription and lexicon transcription. on this data, the system achieved a precision of .89. the decision tree trained on the cmu lexicon performed with a precision of .80 on the same data, showing a .46 reduction in error rate by using the new system. 0.5 0.7 0.9 complete system performance baseline (cmu) 0.801 ear 0.892 combined language score figure 3. performance of the language identification classifier and the language specific letter-to-sound rules combined on 50 english, russian, and arabic words 6. future work and conclusions our results are quite promising. we have shown that we can significantly improve automatic transcription by integrating an n-gram language classifier and language specific transcription decision trees. these promising results warrant further experimentation. there is a wide range of machine learning techniques which we would like to implement in the future as an alternative to using simple n-grams for the first tier of our classifier. we would like to thank fellow researcher dan cer for his suggestion that an unsupervised machine learning approach could be applied to this classification task. preliminary investigations imply that unsupervised ml is promising. 7 lewis et al.: language identification and language specific letter-to-sound rules published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 it is likely that we will see further improvement in the performance of our present day classifier by improving the quality of our training data. suggestions for such improvements have included the addition of non-name based russian and arabic data, comparable to the english brown corpus training data being used. it may also be possible to improve performance by utilizing a filtered russian-name data set, pared down from the large and potentially noisy set we are currently using. the algorithms used in generating the letter-to-sound rules are those commonly accepted as standards. however, we expect that augmenting the size of our transcribed non-english data would increase the quality of output. this will be one of the first areas we will seek to retune as it requires only getting access to more transliterations, transcribing them into phonemes, and incorporating this new data into our training set. finally, our system could easily be extended to handle other non-english word sets via the implementation of additional language specific lts classifiers, and by adding additional classification groups for categorizing among a different set of languages. the functioning of our system requires that there exist a degree of statistical distinction between languages in order to effectively differentiate between them. future work could easily address the theoretical question of assessing the limits on the number of languages it can handle based upon their phonemic similarity and difference. additionally this would serve to test whether the system is robust enough to distinguish between more closely related languages. references coker, k. church and m. liberman. 1990. 'morphology and rhyming: two powerful alternatives to letter-to-sound rules for speech synthesis.' european speech communication association, conference on speech synthesis. jurafsky, daniel and jim martin. 2000. speech and language processing. prentice hall. pellom, bryan and kadri hacioglu. 2003. sonic: the university of colorado continuous speech recognizer, technical report tr-cslr-2001-01. reynar, jeffrey c. and adwait ratnaparkhi. 1997. 'a maximum entropy approach to identifying sentence boundaries.' in proceedings of the fifth conference on applied natural language processing. russell, stuart and peter norvig. 2002. artificial intelligence: a modern approach. prentice hall. urls goldschmidt, paul. 1996. a dictionary of period russian names. http://www.sca.org/heraldry/paul/. retrieved december 2003. cmu sphinx. http://fife.speech.cs.cmu.edu/sphinx/. retrieved december 2003. 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/6 doi: https://doi.org/10.25810/60mf-fn94 http://www.sca.org/heraldry/paul/ http://fife.speech.cs.cmu.edu/sphinx/ colorado research in linguistics 6-2004 language identification and language specific letter-to-sound rules stephen lewis katie mcgrath jeffrey reuppel recommended citation introduction past/related work language classification training data classification algorithm letter to sound rules training data training algorithm results/analysis phase 1: analysis of the classifier phase 2: analysis of the letter-to-sound rules phase 3: analysis of the complete system future work and conclusions thing & place: the cognitive basis and idiosyncrasy of spatial expression in chinese 1 liu: thing & place published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 3 liu: thing & place published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 5 liu: thing & place published by cu scholar, 1994 6 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 7 liu: thing & place published by cu scholar, 1994 8 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 9 liu: thing & place published by cu scholar, 1994 10 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 11 liu: thing & place published by cu scholar, 1994 12 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 13 liu: thing & place published by cu scholar, 1994 14 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 15 liu: thing & place published by cu scholar, 1994 16 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/6 doi: https://doi.org/10.25810/br9g-sd37 colorado research in linguistics 1994 thing & place: the cognitive basis and idiosyncrasy of spatial expression in chinese ningsheng liu recommended citation tmp.1537749392.pdf.tmodn two bolanci complementizers: ii and na 1 a w ad : t w o bo la nc i c om pl em en tiz er s: ii an d na pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ h c 7e -v 83 7 3 a w ad : t w o bo la nc i c om pl em en tiz er s: ii an d na pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /5 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ h c 7e -v 83 7 5 a w ad : t w o bo la nc i c om pl em en tiz er s: ii an d na pu bl ish ed b y c u s ch ol ar , 1 99 3 colorado research in linguistics 5-1993 two bolanci complementizers: ii and na maher awad recommended citation tmp.1538231774.pdf.p9uii a pilot study of the effect of grammatical voice on story recall 1 lenell: a pilot study of the effect of grammatical voice on story recall published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/5 doi: https://doi.org/10.25810/r602-0y11 3 lenell: a pilot study of the effect of grammatical voice on story recall published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/5 doi: https://doi.org/10.25810/r602-0y11 5 lenell: a pilot study of the effect of grammatical voice on story recall published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/5 doi: https://doi.org/10.25810/r602-0y11 colorado research in linguistics 1995 a pilot study of the effect of grammatical voice on story recall elizabeth a. lenell recommended citation tmp.1537647776.pdf.cjchc jemez tones and stress 1 be ll: je m ez t on es a nd s tr es s pu bl ish ed b y c u s ch ol ar , 1 99 3 2 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ vt 8v -3 v4 0 3 be ll: je m ez t on es a nd s tr es s pu bl ish ed b y c u s ch ol ar , 1 99 3 4 co lo ra do r es ea rc h in l in gu ist ic s, v ol . 1 2 [1 99 3] ht tp s:/ /s ch ol ar .co lo ra do .e du /c ril /v ol 12 /is s1 /2 d o i: ht tp s:/ /d oi .o rg /1 0. 25 81 0/ vt 8v -3 v4 0 5 be ll: je m ez t on es a nd s tr es s pu bl ish ed b y c u s ch ol ar , 1 99 3 colorado research in linguistics 5-1993 jemez tones and stress alan bell recommended citation tmp.1538231509.pdf.rdhyr impersonal haber and the road to copula function 1 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 2 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/4 doi: https://doi.org/10.25810/tcz4-fm49 3 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 4 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/4 doi: https://doi.org/10.25810/tcz4-fm49 5 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 6 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/4 doi: https://doi.org/10.25810/tcz4-fm49 7 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 8 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/4 doi: https://doi.org/10.25810/tcz4-fm49 9 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 10 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/4 doi: https://doi.org/10.25810/tcz4-fm49 11 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 12 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/4 doi: https://doi.org/10.25810/tcz4-fm49 13 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 14 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/4 doi: https://doi.org/10.25810/tcz4-fm49 15 nicita: impersonal haber and the road to copula function published by cu scholar, 1997 colorado research in linguistics 1997 impersonal haber and the road to copula function linda m. nicita recommended citation tmp.1537638945.pdf.yofol a queer revolution: reconceptualizing the debate over linguistic reclamation colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by robin brontsema. a queer revolution: reconceptualizing the debate over linguistic reclamation robin brontsema university of colorado at boulder the debate over linguistic reclamation, the appropriation of a pejorative epithet by its target(s), is generally conceived of as a simple binary of support and opposition. i offer an alternative conceptualization that shows both the complex contrasts and commonalities within the debate. specifically, i identify three perspectives: (1) that the term is inseparable from its pejoration and therefore its reclamation is opposed; (2) that it is separable from its pejoration and therefore its reclamation is supported; and (3) that it is inseparable from its pejoration and therefore its reclamation is supported. additionally, by examining different goals within and across reclamations, i demonstrate the difficulty of assigning a fixed outcome of success or failure. although the term queer serves as the primary case study, the terms black, nigger, cunt, and dyke supplement and expand the discussion from a specific study of queer to linguistic reclamation in general. 1. introduction hate speech intended to disable its target simultaneously enables its very resistance; its injurious power is the same fuel that feeds the fire of its counter-appropriation. laying claim to the forbidden, the word as weapon is taken up and taken back by those it seeks to shackle—a self-emancipation that defies hegemonic linguistic ownership and the (ab)use of power. linguistic reclamation, also known as linguistic resignification or reappropriation, refers to the appropriation of a pejorative epithet by its target(s). the linguist melinda yuen-ching chen offers the following definition: “the term ‘reclaiming’ refers to an array of theoretical and conventional interpretations of both linguistic and non-linguistic collective acts in which a derogatory sign or signifier is consciously employed by the ‘original’ target of the derogation, often in a positive or oppositional sense” (1998:130). at the heart of linguistic reclamation is the right of self-definition, of forging and naming one’s own existence. because this self-definition is formed not in one’s own terms but those of another, because it necessarily depends upon the word’s pejoration for its revolutionary resignification, it is never without contestation or controversy. while the controversy over reclamation is generally reduced to a simple binary of support and opposition, i present an alternative conceptualization that accurately represents both the complex contrasts and commonalities within the debate. additionally, by examining different goals within and across reclamations, i demonstrate the difficulty of assigning a fixed outcome of success or failure. although queer is the primary case study, the terms black, nigger, cunt, and dyke supplement and expand the discussion from a specific study of queer to linguistic reclamation in general. 1 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 2. origins of queer 2.1. non-sexual senses the second edition of the oxford english dictionary (simpson and weiner 1989; henceforth oed) identifies queer’s origin as the middle high german twer, signifying ‘cross’ or ‘oblique,’ and provides several definitions, including the following1: adjective: 1a. strange, odd, peculiar, eccentric, in appearance or character. also, of questionable character, suspicious, dubious. 1b. of a person (usually a man): homosexual. hence, of things: pertaining to homosexuals or homosexuality. (united states origin) noun: a (usually male) homosexual. also in combinations, as queer-bashing, the attacking of homosexuals; hence queer-basher. (simpson and weiner 1989: 1014) queer’s original significations did not denote non-normative sexualities, but rather a general non-normativity separable from sexuality. only later in its history would sexuality become the overriding denotation. queer, then, initially could refer to strange objects, places, experiences, persons, etc. without sexual connotations, as in the following literary examples taken from the oed: 1) “the emperor is in that quer case, that he is not able to bid battle” (yonge’s diary of 1621) 2) “i have heard of many queer pranks among my bedfordshire neighbours” (richardon’s pamela of 1742) 3) “it was a queer fancy...but he was a queer subject altogether” (dicken’s barnaby rudge of 1840) (simpson and weiner 1989: 1014) 2.2. sexual senses queer eventually did become associated almost exclusively with non-normative sexuality, an association which has persisted to the present. in contrast to its contemporary usage among queer theorists and self-identified queers (yet similar to its usage in the mass media), by the early 20th century, queer as sexually non-normative was restricted almost exclusively to male homosexual practices, as in the following example from the u.s. children’s bureau’s practical value of scientific study of juvenile delinquents of 1922: “a young man, easily ascertainable to be unusually fine in other characteristics, is probably ‘queer’ in sex tendency” (simspon and weiner 1989: 1014). as george chauncey demonstrates in an examination of terms of self-reference of male homosexuals in new york prior to the second world war, queer co-existed with fairy in the 1910s and 1920s to refer to “homosexuals” (1994: 15-16). far from being synonyms, however, they carried extremely different in-group connotations. differing from queers in their deviant gender status, fairies referred to effeminate, flamboyant males sexually involved with other men. queers, in contrast, were more masculine men 1. for greater clarity, i have condensed the definitions and changed the original formatting. 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 3 who were sexually involved with other men and who generally shunned, even detested, the woman-like behavior of fairies. “the men who identified themselves as part of a distinct category of men primarily on the basis of their homosexual interest rather than their womanlike gender status usually called themselves queer” (chauncey 1994: 16). furthermore, the fairy-queer distinction was not based solely on gender, but on class as well: most queers were men from the middle class who potentially risked more in their professional lives were they to display the femininity typical of fairies (chauncey 1994: 106). because the effeminate fairies’ gender deviance was highly marked and visible, they served as the stereotypical representation of all homosexual men, although there were probably more masculine homosexuals passing as their heterosexual counterparts. heterosexuals used queer and fairy interchangeably and without distinction, thereby homogenizing all men who engaged in sexual activity with other men, regardless of their degree of femininity/masculinity or self-identification (chauncey 1994: 15). homosexuals’ well-defined system of gender classification and the significant differences between queer and fairy were left unrecognized, lost in a classification divided solely along lines of the sex of the partner chosen. queers and fairies were forcibly fused into the same category, one which, because of the latter’s higher degree of visibility, equated homosexuality with femininity. 3. shifting in-group terms 3.1. from queer to gay despite queers’ distance from the femininity associated with fairies, most eventually adopted the term gay, the original territory of the same effeminate men from whom they wished to distance themselves. according to chauncey, “originally referring simply to things pleasurable, by the seventeenth century gay had come to refer more specifically to a life of immoral pleasures and dissipation...a meaning that the ‘faggots’ could easily have drawn on to refer to the homosexual life” (1994: 17). this homosexual reference began with the fairies in the 1920s, employing gay as a code word to be used and understood by homosexuals. a safe word, gay originally denoted lighthearted pleasantness, yet was given a double meaning when used by homosexuals. only those familiar with this specific use of gay would understand it, and therefore, there initially was very little risk in using it with men whose sexuality was unknown. although gay was first used by fairies, queers eventually adopted it since, as a code word, it was understood by all homosexual men: “[gay’s] use by the ‘flaming faggots’ (or ‘fairies’)...led to its adoption as a code word by ‘queers’ who rejected the effeminacy and overtness of the fairy but nonetheless identified themselves as homosexual” (chauncey 1994: 18). however, many queers continued to associate the term with the overt flamboyancy of the fairies—precisely what distanced the two—and therefore rejected the term in spite of its growing popularity among both homosexuals and heterosexuals. according to chauncey, by the second world war gay eventually did replace queer, the latter viewed (especially by younger homosexuals) as derogatory, a pejorative label forced upon them that defined their homosexual interest as deviant, abnormal, and perverse (1994: 19). because of the changing connotations of gay, these younger men 3 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 who embraced the term did not associate it with effeminacy or flamboyance as did the queers. gay grouped all men sexually involved with other men into the same homogenous group; as such, gay, like the out-group usage of queer only a few decades earlier, ignored important differences among those men, coercively forging a common identity based solely upon their sexual object choice and completely disregarding the significance of gender in their self-classification. 3.2. reclamation of queer although gay did overtake queer as the primary label of self-identification among (mainly male) homosexuals, queer experienced a rebirth in the early 1990s due to several factors: the limitations of gay and lesbian as universal categories and homosexuality itself as their foundation; the aids crisis and its behavior-based prevention education and identity-transcendent activism; and queer nation’s coalitional politics of difference and its impact on the reconceptualization of sexual identity. the first instance of queer’s public reclamation came from queer nation, an offspring of the aids activist group aids coalition to unleash power (act-up). queer nation was originally formed in 1990 in new york as a discussion group by several act-up activists discontent with homophobia in aids activism and the invisibility of gays and lesbians within the movement (fraser 1996: 32). the group, originally comprised of members of act-up, soon moved from discussion to the confrontational, direct, and action-oriented activism modeled after act-up. this new coalition chose “queer nation” as its name because of its confrontational nature and marked distance from gay and lesbian. for a coalition committed to fighting homophobia and “queerbashing” through confrontation, queer, “the most popular vernacular term of abuse for homosexuals,” was certainly an appropriate—perhaps perfect—choice (dynes 1990: 1091). rather than being a sign of internalized homophobia, queer highlights homophobia in order to fight it: “[queer] is a way of reminding us how we are perceived by the rest of the world” (kaplan 1992: 36). to take up queer is at once to recognize and revolt against homophobia. queer also served to mark distance from the alleged exclusionary and assimilationist gay and lesbian. an essentializing category of identity will necessarily exclude, and those excluded will in turn contest the categories as universal. the queer of queer nation emphasized the inclusiveness that the more traditional gay and lesbian were seen to lack, advancing beyond their restrictive limits of gender and sexuality to include anything outside of the guarded realm of normalcy, any disruption of the male/female and heterosexual/homosexual binaries. instead of solely relying upon sexual object choice as the basis of sexual identity, queer allowed—and welcomed—a multiplicity of sexualities and genders. difference was not a challenge, but an invitation. queers also publicly rejected the assimilationist tactics of gays and lesbians. refusing to forge their existence within the heterosexual-homosexual polarity, queers chose to wage their war outside of the system. the goal was not to win heterosexual support or approval; therefore, their battle did not model a civil rights movement, struggling for equal rights for an oppressed minority (duggan 1992: 16). queers associated gay and lesbian with an unquestioning acceptance of the status quo and an essentializing understanding of sexuality and gender. queer, in contrast, was associated with a radical, confrontational challenge to the status quo, and a constructionist 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 5 understanding of sexuality and gender. queer was by no means intended to be a synonym for gay and lesbian. although they may share a common denotation (but not necessarily), the connotation was different; they were never meant to be interchangeable equals. the differences between queer and gay / lesbian mirror the history of black and negro. during the black power movement beginning in the late 1960s, it was argued by black activists that a negro was passive, docile, acquiescent—s/he did not challenge racism or fight against the status quo; a black person, however, was active, fierce, challenging—s/he rebelled against racism and the status quo (bennett 1967: 47). these significant distinctions, however, were later lost, just as queer (as will be discussed) largely became synonymous with gay. 4. reconceptualizing the debate the reclamation of queer has been largely fragmented, limitedly accepted, and highly contested. queer has been popularly both opposed and supported because of its pejoration. although this debate is commonly perceived as a simple reduction to two opposing sides—those who support the reclamation of queer and those who oppose it— the debate is actually more complex and in order to understand the reclamation of queer and linguistic reclamation in general, an alternative conceptualization is not only useful, but necessary. the traditional representation of the debate over reclaiming a pejorative word is usually a simple binary of support or opposition, as shown in figure 1. a more appropriate representation capturing the diversity of the debate, however, is illustrated in figure 2. each of the three perspectives will in turn be discussed, both their characteristics and limitations. figure 1 traditional representation of the debate over linguistic reclamation figure 2 reconceptualization of the debate over linguistic reclamation reclamation opposed reclamation supported reclamation opposed pejoration inseparable reclamation supported pejoration separable perspective 1 perspective 2perspective 3 5 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 4.1. perspective one—pejoration inseparable: reclamation opposed the first perspective is that queer is inseparable from its pejoration and therefore should not be used. according to this viewpoint, using queer (or any pejorative epithet) can only be self-degrading and disrespectful, a repetition of the intolerance and hate that the word somehow encapsulates and carries with it. the pejoration cannot be removed from the word; indeed, the word and its pejorative meaning are indistinguishable. the hate, the pain, the violence is locked in that word forever, and therefore the word itself must be locked away in the attic of a collective linguistic memory. bringing out the word would necessarily bring out the pain. because the word is rendered inseparable from its injurious power, many who oppose its reclamation are those who have directly suffered from its infliction and still bear the scars that can never completely heal: “we still lick the psychic and physical wounds inflicted by the word ‘queer’” (sillanpoa 1994: 57). this naturally results in an agebased division between supporters and opponents of the word’s reclamation, with those of an older generation who have experienced it as abusive and violent opposing its ingroup circulation, which is often viewed in terms of the younger generation’s arrogance and disrespect. pain marks the boundary of an uncloseable gap between generations. according to this first perspective, because pejoration is external, coming from without and not from within—that is, the term is used as hate speech by the out-group (heterosexuals) against the in-group—it can never be reclaimed as one’s own. linguistic ownership is permanent: the stronghold is unwavering, a frozen grip that will never loosen. queer is and will always be an outsider’s weapon that forcefully establishes bounds of legitimacy; attempting to take back this weapon is not only futile but selfdefeating. the master’s house cannot be destroyed using the master’s tools (lourde 1984: 113). this particular viewpoint is clearly not limited to the reclamation of queer, but can be found in any debate over reclamation in general. a personal online journal discussing the author’s opposition to the reclamation of cunt, for example, expresses this common viewpoint: "you end up joining that same force of oppression you’re trying (question mark) to work against….self love…isn’t going to be found by their words" (carmel 2003). most noteworthy is the author’s consideration of linguistic ownership as fixed, unable to be transferred, transformed, nullified. if ownership is fixed, then it naturally follows that reclamation is at best naively optimistic, at worst inevitably impossible. as permanent property of those who use the word as hate speech, the use of queer among homosexuals can only reinforce a homophobic society’s common stereotypes and fears. the so-called reclamation is not a revolution—it is a repetition. there is a prevalent anxiety that the in-group use of queer “only serves to fuel existing prejudice and may even lead to an increase in discrimination and violence” (watney qtd. in jagose 1996: 106). lying under this anxiety is the assumption that the out-group would conflate the two distinct uses of queer into one, rendering both equally anti-gay. that the outgroup would fail to recognize the distinct, new use of queer among homosexuals is taken for granted. 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 7 4.1.1. problems with this view those who would claim that queer “has always been, is now and will always be an insulting, homophobic epithet” (saunders qtd. in thomas 1995: 76) fail to recognize the nature of language, the constant change of words—their births, deaths, resurrections, metamorphoses. new words will be created, old ones will die, old words will take on new meanings, new words will take on old meanings: language is dynamic and everchanging. change is the only constant. the first perspective not only freezes meaning in time, but linguistic ownership as well: meaning and control over its production are misunderstood to be fixed and stable. words, however, are not exclusively owned or used. one usage does not disallow others; one group’s pejorative use of a word does not prevent another group—indeed, its targets—from using it in new contexts and with differing intentions. perhaps the only way to mitigate the injurious power of the word is for its very target to take it up as its own. that linguistic ownership is unfixed and unstable can be illustrated with nigger (or nigga). nigger derives from the portuguese negro, translated as black, to refer to african slaves and was later adopted by the british and americans. geneva smitherman states that “'negro’ and ‘nigger’ were used interchangeably and without any apparent distinction....it was not until the twentieth century that whites began to semantically distinguish ‘negro’ and ‘nigger,’ with the latter term becoming a racial epithet” (1977: 36). those who cannot conceive of nigger as anything but a racial epithet subscribe to an out-group interpretation that fails to recognize the complexity and diversity of nigger's ingroup usage. far from being restricted solely as a derogation that maintains racial subordination, smitherman identifies seven contemporary, in-group uses of nigga (as opposed to nigger, often associated with outgroup usage)2: 1) close friend, backup 2) someone who is culturally black 3) synonym for blacks or african-americans 4) african american women’s term for the black man as lover/partner/significant other 5) rebellious, fearless, unconventional, in-yo-face black man 6) derogatory, similar to pejorative use by out-group 7) any cool, down person who is deeply rooted in hip hop culture (2000: 210-11) the diversity of nigga’s uses and significations is testament to the fluidity and temporality of linguistic ownership. those who sought to shackle a people with a word witnessed its emancipation, its removal from an original subordination to a freedom to take on a multiplicity of meaning, “not in subjection to racial subordination but in defiance of it” (kennedy 2002: 47). 2. i have condensed this list and changed the original formatting. 7 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 4.2. perspective two—pejoration separable: reclamation supported in contrast to the first perspective, that queer cannot be separated from its pejoration and therefore should not be used, this second perspective is that it can indeed be separated and for that reason should be used. according to this perspective, queer can be made neutral or even positive. the goal is for it to lose its stigma so that it can no longer offend or injure. the only way of conquering pain, then, is to work with it—not to ignore or hide from it: “we have to take the q-word back in order for it not to cause pain” (andrew qtd. in thomas 1995: 79). expropriation is a necessary response. the following expresses a similar viewpoint. “several people in the gay community on a quest to end the hate attached to the word took control of the word; they decided to embrace the word. by doing so, they quelled yet another weapon from the homophobic’s arsenal of hate...learn to love the word and the tools of hate are quashed” (garner qtd. in thomas 1995: 79). although this goal of queer’s reclamation is the depletion of its injurious power, the majority of those who support its reclamation are those who have no direct experience with queer as a term of abuse; again, pain marks the generational division. many, especially youth, who self-identify as queer never had the term used against them as a homophobic epithet and therefore can easily use it with pride—there are no traumatic memories embedded in the word, no associations that with its enunciation recall tragedy, fear, hate, and rage. as this second perspective asserts the separability of the epithet from its pejoration, there exist two differing goals of its reclamation: neutralization and value reversal. to neutralize hate speech is to render it ineffective, to nullify its force, to remove its biting sting. a word neutralized is a word deadened, incapable of inspiring shame, or pride. if neutralization is the goal, then the reclamation must necessarily die once its driving fire has turned to ash. the most successful reclamation, then, is that which extinguishes itself. the power of the word to injure is ironically the fuel of the movement that seeks to remove that very power. reclamation feeds on the negative energy of the word; therefore, once this negativity is neutralized, the word reclaimed must necessarily lose that driving, revolutionary force. to become acceptable, it must become less potent. paradoxically, a successful reclamation reaches its death at the same time it reaches its goal. slowly, quietly, almost unnoticeably, it dies, all but forgotten. in contrast to neutralization, value reversal is the transformation of a negative value into a positive one. to reverse a word’s value is to completely turn it around 180 degrees to its opposite, to steal it from its injurious trajectory to send it on the opposite path—to reverse value is to exchange opposites. value reversal is largely assumed to be the goal of linguistic reclamation. certainly, it is a—even if not the only—goal of many movements to reclaim a word. this value reversal, for example, is at the heart of the contemporary feminist movement to reclaim cunt. today cunt could be considered the most abusive, misogynist epithet used against women, derogatively signifying not only female genitalia but women in general. although it is now known as a term of abuse, however, cunt originally had neutral or positive connotations; in fact, many goddess’ names share this root (hunt). the oed provides several examples of cunt: 1) "gropecuntelane" (a london street name in ekwall’s street names of city of london of 1230) 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 9 2) “for ilka hair upon her c—t, was worth a royal ransom” (burns' merry muses of 1800) 3) “what’s the cunt want to come down ‘ere buggering us about for, ‘aven’t we done enough bloody work in th’ week?” manning's middle parts of fortune of 1929) (simpson and weiner 1989: 130) over time, cunt clearly acquired extremely negative connotations, becoming an unrivaled misogynist epithet. feminists have called for a collective reclamation of cunt, as does igna muscio in cunt: a declaration of independence, in which she clearly articulates her goal as value reversal: “when viewed as a positive force in the language of women...the negative power of ‘cunt’ falls in upon itself” (2002: xxvi) (italics added). for muscio and many others, to reclaim cunt is to reverse its value, to replace its negative connotative value with a positive one. this value reversal channels the power that the word already contains, tapping this source of energy in order to create its very opposite. it is nothing less than a revolutionary reversal of opposites. in addition to its goals of neutralization or value reversal, this second perspective differs from the first in its understanding of linguistic ownership. unlike the first perspective that regards ownership as fixed and permanent, this perspective stresses its transferability, the taking-back in order to make it one’s own. that the origin of queer’s pejoration is external does not doom it to slavery, assigning it a tragic and unalterable fate. it is this defeating faith in fate against which reclamation rebels. the destiny of a word can be steered off its track, turned around, and sent on a new path in a different— perhaps opposite—direction. muscio demonstrates her faith in the transferability of ownership, demanding that cunt be seized, rescued from its misogyny: “i posit that we’re free to seize a word that was kidnapped and co-opted in a pain-filled, distant past, with a ransom that cost our grandmothers’ freedom, children, traditions, pride and land” (2002: 9). if a word could be kidnapped, it could also be taken back. 4.2.1. problems with this view because the pejorative power of the word fuels the very movement to deplete this power, those who would assert that the word could be separated from this pejoration fail to recognize its complexity. a reclaimed word partly depends on the pejoration that drove its reclamation; indeed, this very pejoration allowed for its metamorphosis, its rebirth into something new and different from its original derogation, but never completely separate from it. perhaps the strongest criticism of this view is that its supporters are in a continuous state of reactionary response. if x is used against a group and is reclaimed, what about y? and what about z? if the word that wounds must be appropriated in order for its injurious trajectory to come to an end, then it follows that all such words must be reclaimed so that they no longer carry power. other pejorative terms will always co-exist with the term being reclaimed. as long as there is homophobia (sexism, racism, etc.), language will always express it. similarly, homophobia cannot be depleted with the reclamation of queer: it is symptomatic of a social disease that cannot be cured with one single change. unlike the reclaimed queer, a self-created term of identification steps out of this state of response, as is the case of african american. although jesse jackson is usually 9 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 credited for african american, he only brought to public attention what had been coined in 1988 by dr. ramona h. edelin, then president of the national urban coalition (smitherman 1991: 115). most american slave descendants (asd) (baugh 1991: 133) who use african american over black prefer the former’s explicit identification with africa, a dual heritage not expressed by the latter (smitherman 1991: 125). furthermore, many asd feel that black as a color label is inadequate: unlike ‘whites,’ who may further identify as european-americans or irish-americans, for example, black fuses color, ethnicity, and nationality into one. african american received phenomenal national public attention when jackson made a public call for its adoption in 1988: “'just as we were called colored, but were not that, and then negro, but not that, to be called black is just as baseless. every ethnic group in this country has reference to some cultural base. african americans have hit that level of maturity'” (qtd. in baugh 1991: 133). fifteen years after jackson’s call, african american has certainly become the most acceptable racial designation for asd and can be heard and seen throughout academia, government, the media, and also on the streets. african american has become the most politically correct term, safe and inoffensive. its lack of offensiveness is due in large part to the fact that asd themselves created the term: the signified created its signification. its origin is internal; there is no pejoration, no hate to overcome, to twist, to rework into something new. hate was never at its heart; therefore, it must not rely upon racism and hate in order to fuel its revolution—it transcends it. instead of being trapped in a continuous cycle of reaction, it acts upon itself for itself, choosing not to step into that trap at all. 4.3. perspective three—pejoration inseparable: reclamation supported sharing the assertion of the first perspective that queer is inseparable from its pejoration, proponents of the third perspective base their support of the reclamation upon this very inseparability, in stark contrast to the first’s opposition. this perspective shows a third identifiable goal of reclamation (in addition to neutralization and value reversal): stigma exploitation. according to this perspective, queer should be reclaimed by its original targets and purposefully retain its stigma—a confrontational, revolutionary call. instead of erasing the stigma, it seeks to highlight it: “[the] reclaiming of pejorative terms…does not conspire to remove the derogation, or even to undo it, but recast it into a sign of a stigma, rather than a tool of a stigma” (chen 1998: 138). instead of being a selfdefeating, homophobic statement of one’s abnormality or ‘queerness,’ queer boldly questions the very construct of sexual abnormality (and thus normality). as butler argues, “within the very signification that is ‘queer,’ we read a resignifying practice in which the desanctioning power of the name ‘queer’ is reversed to sanction a contestation of the terms of sexual legitimacy” (1993: 232). to declare oneself queer is to question the social construction and regulation of sexual normalcy. in contrast to the first perspective, that more or less equates the in-group and out-group usages of queer, this perspective regards them as very much distinct. indeed, in-group usage radically differs from out-group usage by questioning the very assumptions of normality upon which its pejoration is based. in this sense, then, queer absolutely needs its stigma in order to confront the construction of its abnormality. to remove its pejoration is to bring the revolution to its end. there are those optimistic supporters of queer’s pejorative power who believe that not only should queer retain its stigma, but that it must necessarily do so, that it is 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 11 impossible for it to lose this power. this “immunity to domestication” (jagose 1996: 106) is optimistically believed to protect queer from neutralization, a mainstreaming that would dilute its potency to the point of undetectability. queer will forever retain its stigma because it consciously chooses it, fighting against those who seek to catch and freeze it, fixing it to an unalterable fate. intent is the omnipotent force that guides queer away from a fatal fixedness and ensures the success of its battle against death. 4.3.1. problems with this view the optimism of this perspective unfortunately puts too much faith in will and intent: they alone cannot control the fate of a word. ironically, reclamation is testament to the inherent frailty of intent: just as those who used queer pejoratively could not know that the word would be taken up and twisted by its very targets, optimistic supporters of queer deny that the word could be used in ways they never intended. ironically, those with blind faith in the reclamation of queer fail to learn its lesson, the essence of any reclamation: that intent can—and indeed, sometimes must—be betrayed. consciously choosing queer’s stigma does not guarantee its permanence; the future of a reclaimed word cannot be determined in advance. the history of black shows that revolutionary intent does not pre-determine the future of a word, that intent can be betrayed even when a word is said to be “reclaimed.” although most associate the birth of black with stokeley carmichael and the black power movement of the late 1960s, it was actually employed as a term of self-reference, along with the more common african, at least a century before the emancipation proclamation of 1863 (bennett 1967: 48). colored and negro (and eventually the capitalized negro), however, became the most popular terms of self-reference and it was not until 1966 that the activist carmichael made a national call for “black power,” demanding that black (note capitalization) replace negro (smitherman 1991: 121). as previously mentioned, negro was considered to be a “slave-oriented epithet which was imposed on americans of african descent by slavemasters,” connoting docility, passivity, and meek acceptance of the status quo that did nothing to challenge it (bennett 1967: 54). in contrast, black was for "'black brothers and sisters who are emancipating themselves'" (bennett 1967: 47). black was a confrontational “repudiation of whiteness and the rejection of assimilation” that sought to revalue that which was so intensely despised: blackness (smitherman 1991: 121). this revolutionary revaluation was intended for blacks by blacks, yet “when whites became familiarized with the term, they perceived that this was an unobjectionable way to talk to blacks about blacks, and with this perception the nuances of the black inversion were unrecognized….when whites use…black, they just substitute…negro” (holt 1972: 158). the original energy of black was betrayed and subsequently died as it was not used with the same vital radicalism. instead of forcing racists to confront their hatred and speak it out loud, their racism was simply given a new mask to wear. once black was mainstreamed, it was also doomed— perhaps a necessary outcome of a dubiously successful reclamation. the issue of queer’s (in)separability from its pejoration clearly marks dividing lines in the debate over reclamation. can a word, as the second perspective suggests, completely shed its history of abuse? can it experience a transformational metamorphosis so that it no longer resembles that which it once was? can a medicine be concocted from its poison? can its fate be forcefully determined? or, as the first and 11 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 third perspectives suggest, does the word forever carry hate within it? is it subservient to its own history? must it re-enact the drama of its genesis as hate speech? can it never remove its shackles and declare it own emancipation? because the energy of hate speech is the very fuel of its re-appropriation, a word reclaimed may be seen as necessarily both inseparable and separable from its pejoration. its use as hate speech not only influences its reclamation, but is its cause, its origin, its driving fire. even if a word could be completely reclaimed, its former abuses and injuries are necessarily connected to its present usage; indeed, it was only because of that former pejorative meaning that its new positive or neutral meaning could be created and understood. furthermore, a word’s fate can never be determined in advance; a requisite uncertainty weakens faith in the omnipotence of reclamation. ironically, a so-called successful reclamation proves that ownership is not fixed and that fate is not predetermined, demonstrating its success as it simultaneously confirms its own temporality. 5. several uses of queer coexisting far from being limited solely to positive in-group use and negative out-group use, several uses of queer co-exist; whether competing with each other or living together harmoniously, they show not the success of queer’s reclamation, but the myriad of possibilities it has created. 5.1. self-identified queers self-identified queers familiar with queer theory do not employ the term, as commonly believed, as a simple replacement for gay or lesbian. rather, it serves as a conscious contestation of those very terms: by highlighting the bounds of legitimacy, queer simultaneously contests them. furthermore, queerness is not based, in stark contrast to gay and lesbian, on sexual object choice, and as such, is not limited to or by same-sex desire. its inherent inclusiveness allows among its ranks not only queer gays, lesbians, bisexuals, and transgendered, but also queer straights, sadomasochists, fetishists, etc.—any non-normative sexuality or sexual practice could theoretically claim queerness. this use of queer by those who self-identify as such has not extended to the outgroup, often using it as a synonym of gay and lesbian or simply gay in its genderexclusive sense. although it has been revalued by those who claim it as an identity (or a kind of anti-identity) label, the larger society has generally failed to recognize its nuances. in this sense, it mirrors the history of black, intended as a radical revaluation by the in-group yet nevertheless regarded by the out-group as merely a less-offensive synonym of negro. the histories of both black and queer show that intent is not sufficient: it may be misunderstood, ignored, or betrayed. 5.2. non-queer gays and lesbians those who do identify as gay or lesbian but not as queer use it as a convenient catchall term that efficiently encompasses the wordy “gay, lesbian, bisexual, and transgendered.” synonymous with gay and lesbian, and occasionally bisexual, and even more rarely transgendered, this use betrays the radical intent of self-identified queers and queer theorists, and can be found throughout traditional gay and lesbian studies, 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 13 anthologies, conferences, events, etc. tragically yet perhaps predictably, queer is equated with the very terms against which it rebels. while often employed as an efficient substitute in the lgbt community, the success of queer’s alleged inclusiveness is questionable. in theory, spanning over lgbt, it is meant to include men, women, and transgendered; male homosexuals, female homosexuals, and bisexuals. in practice, however, queer often refers to gay men and lesbians, disregarding bisexuals as well as transgendered. although it may not be as sexually inclusive as it claims to be, its use in the lgbt community does show it as largely gender-inclusive—an inclusiveness, however, that is ignored by the media and much of the public. 5.3. popular television television can be accredited with the widespread proliferation of yet another use of queer—a trendy, hip replacement of gay, yet faithfully continuing its gender-exclusive tradition. despite the claim that queer is gender-neutral, its use in popular television clearly associates it with male homosexuality. the television program queer as folk focuses primarily on the lives of gay men, and queer eye for the straight guy has five pairs of “queer” eyes—all belonging to gay men. there is nothing necessarily queer about these gay men: same-sex desire does not exclusively indicate queerness, nor does queerness exclusively indicate homosexuality. however, as employed today in television, gay men are still those signified by queer, albeit with much less offense and much more playful trendiness. although popular television has certainly made queer more acceptable, it has done so in ways that have betrayed its usage by self-identified queers, queer theorists, and gays and lesbians. because it is used as a hip synonym of gay, it loses the radicalism with which self-identified queers and queer theorists use the term—they never intended it as a simple replacement for an out-dated term. in addition, because it is used to refer mainly to gay men and not women, it loses the gender inclusiveness with which many selfidentified gays and lesbians use the term. popular television has succeeded in making queer more acceptable, but who has been excepted as a result? 5.4. hate speech the resurgence of queer—by those who self-identify as such, by non-queer gays and lesbians, and by popular television—has not only not eliminated pejorative use of the term, but may actually have raised it from the dead of a linguistic memory, bringing back to life its injurious power. throughout its “reclamation,” queer has continued to be used pejoratively. however, it had also lost linguistic currency, especially among those of a younger generation. the gender-specific fag and dyke replaced queer as the most popular homophobic epithet, and some of a younger generation had little or no familiarity with queer. it was lost (or hidden) from collective linguistic memory—it no longer had any currency and therefore neither offended nor rebelled: its power was reduced to none. the resurgence of queer beginning in the early 1990s, then, brought it back to linguistic life and restored its relevance. this made possible not only its reclamation, but also its re-pejoration. whereas for some time it was lost from cultural memory, or at least no longer at the forefront of consciousness, it is contemporarily used once again as hate 13 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 14 speech. ironically, the very movement to reclaim queer could actually have restored its injurious power for some, bringing back to life what was dead and forgotten. the re-pejoration of queer, however, is not a guarantee of its efficacy: hate speech can only succeed as such if the target chooses to be its victim. if the target chooses, however, to take up the word as its own, hate speech must fail—the intent to injure does not guarantee that it can nor will. it is necessarily on shaky ground, uncertain of its own efficacy. one cannot forcefully be named that which has been self-chosen; the name then cannot pierce and puncture, but can only be sent back, dull and deadened. ironically, the success of hate speech is determined by the very person made to be deprived of power: the target can choose to take in the word, or to take it back. the co-existence of several uses of queer testifies not to its failure, nor its success, but to the power of reclamation to bring to queer a range of new possibilities, and a future unknown. queer's reclamation has unleashed it from its history of derogation, freeing it to explore new possibilities in a future that cannot be predetermined nor predicted. it is unknown if its pejorative use will cease as different homophobic epithets replace it; it is unknown if its use in popular television will supersede all others; it is unknown if it will continue to be employed as an umbrella term in the lgbt community; queer’s destination is unknown, as it should be. 6. conclusion: successes and failures to assign to reclamation one of two outcomes–success or failure–is to drastically reduce its complexity to an impossibly simple choice between two extremes, when it can rarely fully occupy either. this indeterminacy can be explained in terms of reclamation itself as a process, the ambiguous appearance of success or failure, and the different goals both across and within reclamations. linguistic reclamation is always a process without a clearly marked end. there is no specific final destination that can be found where all pejorative use of a word has ceded to positive use, where success and failure can be captured, measured, weighed. language is constantly changing and meaning is constantly (re)constructed—its nature is not static. to restrict reclamation, then, to one delineative end is to betray something essential of language itself, and is therefore necessarily inaccurate and incomplete. furthermore, the appearance of success or failure may be highly ambiguous and misleading. this ambiguity is perhaps best illustrated with dyke: although dyke continues to be used pejoratively, it is often used positively, with pride, by the in-group. indeed, because of its very pejoration, dyke claims a political fierceness and anti-assimilationism that lesbian lacks, the latter seen to appeal to male, heterosexual, white, middle-class taste. again, although they may share a common denotation, the connotations are extremely different. has dyke failed as a reclaimed word since the out-group continues to use it as hate speech? or does its in-group use alone testify to its success? the appearance of success or failure is additionally ambiguous and untrustworthy because a once pejorative word now mainstreamed is not necessarily a word reclaimed. as discussed, black was popularly and publicly used in a way that was never willed or intended by those who sought a revaluation of blackness. is its common and inoffensive use, then, sign of its success? likewise, is the proliferation of queer as a synonym for the gender-exclusive gay proof that its reclamation has succeeded? 14 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 15 in order to classify a reclamation as a success or a failure, the relationship between goal and outcome must be considered. if a reclamation has succeeded, there would be a match between goal and outcome; in contrast, if it has failed, there would be a mis-match, an incongruency between the two. if to measure the success of a reclamation is to examine the connection between goal and outcome, then it cannot be assumed that there is one common, unifying goal of any and all reclamations—goals differ across and within reclamations and from perspective to perspective. the three perspectives shown in the reconceptualization of the debate over linguistic reclamation (figure 2) deem not only what is desirable to achieve, but also what is possible. to assert, as do those sharing the first perspective, that a word will always have the power to injure is to declare the futility of reclamation. the goal of the opposition to reclamation, then, is censorship, whether of those who use the word to wound or those who use it to heal. if, however, this same power to injure is viewed as the power to transform, as in the second perspective, then reclamation can conquer hate speech by making it ineffective or by reversing its value. those sharing the third perspective believe in this transformational energy of hate speech; however, they also believe that a word's history of injury is inescapable. therefore, reclamation can conquer hate speech not by erasing its stigma but by exploiting it, by using that same power to achieve different ends. as discussed, there are at least three identifiable goals of reclamation, which can be classified as the following: 1) value reversal 2) neutralization 3) stigma exploitation because goals differ across and within reclamations, determining success is highly problematic. if the goal is value reversal, would the word's mainstream use as a simple, synonymous substitute for an outdated term be proof of its unshaken success? if neutralization is the goal, would effective political rallying under a pejorative epithet mark its success? likewise, if stigma exploitation is the goal, would its politicallycorrect, inoffensive usage render it successful? in addition, goals differ in terms of in-group and out-group usage. it may be desired that the reclamation extend to and include the out-group, those who were not and are not targets of the word as hate speech. in contrast, it may be that only in-group reclamation is desired—taking back and staking claim, marking with signs of trespass who has the right granted and who has the right denied. this is often regarded as the case with nigger, that, while it may be acceptable for asd to use it freely, it is off-limits to whites, whose usage of nigger cannot be the same, given its history and the general history of racial oppression and racial relations in the united states. the filmmaker spike lee articulated this view when criticizing fellow filmmaker quentin tarantino’s use of nigger. when it was mentioned that lee himself often uses nigger in his films, he replied that he had "'more of a right to use [the n-word]'" (qtd. in kennedy 2002: 131). to measure success, then, in-group and out-group usage must also be considered. if the goal is solely in-group usage, would its out-group adoption and following testify to its success—or failure? likewise, if the goal is in-group as well as out-group usage, would it still be considered successful if the usages radically differed, if out-group usage 15 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 16 betrayed the very intent of the reclamation? indeed, for a reclamation to extend to those who never suffered from the word as wound, must it necessarily lose its vital radicalism? is it an inevitable outcome? must a reclamation aimed to eliminate or reverse an epithet’s homophobia, misogyny, racism, etc. unavoidably suffer a tragic misinterpretation and further resignification in a homophobic, misogynist, racist society? the sociolinguists susan ehrlich and ruth king assert that linguistic meanings are, to a large extent, determined by the dominant culture’s social values and attitudes – i.e., they are socially constructed and constituted; hence terms initially introduced to be nonsexist, nonracist, or even feminist may...lose their intended meanings in the mouths and ears of a sexist, racist speech community and culture. (1994: 60) the success of a movement, then, to remove, reverse, or rework an epithet’s derogation for both the in-group and out-group is at least partially dependent upon that against which it fights; its fate is not fully dependent upon nor determined by those who demand a revolutionary resignification. while linguistic reclamation may not produce clear victories, it does prove that the right of self-definition is a worthy cause for revolution. to appropriate the power of naming and reclaim the derogatory name that one never chose nor willed is to rebel against the speech of hate intended to injure. linguistic reclamation is a courageous selfemancipation that boldly moves from a tragic, painful past into a future full of uncertainty, full of doubt—and full of possibility. 16 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/9 doi: https://doi.org/10.25810/dky3-zq57 a queer revolution 17 references baugh, john. 1991. 'the politicization of changing terms of self-reference among american slave descendants'. american speech 66(2): 133-146. bennett, lerone, jr. 1967, november. 'what's in a name?'. ebony: 46-54. butler, judith. 1993. bodies that matter: on the discursive limits of "sex", new york: routledge. carmel. 2003, september 8. ‘carmel’s journal’ . chauncey, george. 1994. gay new york: gender, urban culture, and the makings of the gay male world, 1890-1940, new york: basic books. chen, melinda yuen-ching. 1998. ‘“i am an animal!”: lexical reappropriation, performativity, and queer’. engendering communication: proceedings from the fifth berkeley women and language conference: 128-140. duggan, lisa. 1992. ‘making it perfectly queer’. socialist review 22(1): 11-31. dynes, wayne. 1990. ‘queer’. encyclopedia of homosexuality, new york: garland. ehrlich, susan and ruth king. 1994. ‘feminist meanings and the (de)politicization of the lexicon’. language in society 23: 59-76. fraser, michael. 1996. ‘identity and representation as challenges to social movement theory: a case study of queer nation’. mainstream and margins: cultural politics in the 90s, westport, ct: greenwood. holt, grace sims. 1972. ‘“inversion” in black communication’. rappin' and stylin' out: communication in urban black america: 152-159. hunt, matthew. 2003, september 14. cunt: a cultural history. jagose, annamarie. 1996. queer theory: an introduction, new york: new york up. kaplan, esther. 1992, august 14. ‘a queer manifesto’. village voice: 36. kennedy, randall. 2002. nigger: the strange career of a troublesome word, new york: pantheon. lourde, audre. 1984. sister outsider: essays and speeches, new york: crossing. muscio, inga. 2002. cunt: a declaration of independence, new york: seal. sillanpoa, wally. 1994. ‘please, don’t call us “queer”’. radical teacher 45: 56-57. simpson, j.a. and e.s.c. weiner. 1989. the oxford english dictionary, second edition. oxford: clarendon smitherman, geneva. 1977. talkin and testifyin: the language of black america, boston: houghton mifflin. smitherman, geneva. 1991. ‘what is africa to me?: language, ideology and african american’. american speech 66(2): 115-132. smitherman, geneva. 2000. black talk: words and phrases from the hood to the amen corner, boston: houghton mifflin. thomas, david. 1995. ‘the 'q' word’. socialist review 25(1): 69-93. 17 brontsema: a queer revolution published by cu scholar, 2004 colorado research in linguistics 6-2004 a queer revolution: reconceptualizing the debate over linguistic reclamation robin brontsema recommended citation microsoft word paper_brontsema.doc marginalization of alternative gender and sexual identities: the role of normative discursive practices in chilean society marginalization of alternative gender and sexual identities: the role of normative discursive practices in chilean society sara balder university of colorado in chile, a variety of conventionalized metonymic comments and address terms are used in everyday discursive practices as a means of ridiculing gender and sexual minorities. language provides a tool for associating gender or sexually non-normative males with women, either by alluding to their effeminacy, their sexual passivity, or a combination of the two. by presupposing an intrinsic relationship between gender and sexual orientation, these heterosexist comments play a vital role in maintaining the standard social expectations surrounding gender and sexuality, consequently subordinating individuals who do not adhere to these norms. although seen as harmless jokes by those who regularly employ them, i argue that by derogating gender and sexual minorities, heterosexist commentary is a powerful force that engenders the reproduction of heteronormative beliefs in society. 1. introduction the social construction of gender and sexual identity in chile emerges from a well-established and thriving system of beliefs surrounding acceptable gender and sexual norms. discursive practices such as the normative use of heterosexist comments and address terms play a significant role in creating pressure for members of chilean society to adhere to these rigid social expectations. in this paper, which is a component of a larger project involving language and homophobia in chile, i address the phenomenon of how discursive practices promote the marginalization of homosexual or gender non-normative males. specifically, i focus on discourse samples that are ‘heterosexist’, meaning that they align with “the institutionalized assumption that everyone is heterosexual or should be, and that heterosexuality is inherently superior and preferable to homosexuality or bisexuality” (marrones 2001:26). i would like to show that the conventionalized use of such expressions naturalizes gendernormative heterosexuality, the consequence of which is that gender and sexual minorities are ‘marked’ as abnormal or inferior. the language samples i examine in this paper, which target males that are gender and/or sexually non-normative (i.e. males that are effeminate and/or homosexual), illustrate that the perception of gays as abnormal or inferior stems from the association of non-normative males with women. as in many patriarchal latin american countries, women are perceived in chile to be inferior members of society. by analyzing language samples used in everyday discourse, i show that this association is made primarily in two ways: first, by alluding to colorado research in linguistics. june 2005. vol. 18, issue 1. boulder: university of colorado. © 2005 by sara balder. 1 balder: marginalization of alternative gender and sexual identities published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) effeminate gender expression of gay males, and second, by alluding to their passive sexual role. by using conventionalized verbal insults and address terms to ascribe feminine gender and sexual traits to men, language is used in society to convey the hegemonic norms of dominant culture. the prevalence of the heterosexist commentary i examine reproduces the hetero-normative ideology that governs linguistic and social practices in chile. i approach the concept of identity as a social phenomenon by illustrating that the use of heterosexist commentary facilitates the social positioning of both self and other (see bucholtz & hall forthcoming), whereby speakers index themselves as heteronormative by labeling someone else as non-normative and as such, subordinate. 2. heterosexist commentary: implied inferiority of homosexual males the heterosexist language used in chile presupposes a direct relationship between gender identity and sexual orientation. in other words, all gendernormative individuals (i.e. feminine women and masculine men) are inherently surmised to be heterosexual. consequently, anyone who does not adhere to the gender norms prescribed for their particular sex is labeled as homosexual. by derogating individuals that demonstrate non-normative gender traits or sexual orientation, hetero-normative discursive practices in chile exemplify the importance chileans place on asserting heterosexuality as part of their normative gender. this phenomenon can be observed in a number of typical chilean verbal comments. during my research period in chile, i collected a total of twelve conventionalized heterosexist verbal insults, which i categorized into three distinct but related groups: allusion to gender non-normativity: se le da vuelta el paraguas. ‘his umbrella gets inverted’ se le queda la pata atrás. ‘his foot gets left behind’ se le quema el arroz. ‘his rice is getting burnt’ se le apaga el calefón. ‘his pilot light goes out’ reference to effeminate physical gestures: se le cae el completo. ‘he drops his hotdog’ el tiene maletas imaginarias. ‘he has imaginary suitcases’ allusion to homosexual acts or homoerotic desire: el hace géminis sesenta y nueve. ‘he does gemini sixty-nine’ el muerde la almohada. ‘he bites the pillow’ se le chorrea el completo. ‘his hotdog is dripping’ le gustan las tunas. ‘he likes cactus fruit’ le gusta por el camino de tierra. ‘he likes to take the dirt road’ le gusta por detroit. ‘he likes it by ‘detroit’’ 2 2 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/3 doi: https://doi.org/10.25810/j1vk-ck75 marginalization of alternative gender and sexual identities these comments demonstrate the overlapping conceptualization of gender and sexual orientation in chile, in that male effeminacy is perceived as inseparable from male homosexuality. the first two categories of phrases rely on allusion to effeminate gender expression to associate gay males with women; the third category of phrases relies on reference to attraction to other men or taking the passive sexual role in order to make this association. i will look at five comments that allude to non-normative gender characteristics of the referent: se le da vuelta el paraguas ‘his umbrella gets inverted’, se le queda la pata atrás ‘his foot gets left behind’, se le quema el arroz ‘his rice is getting burnt’, se le cae el completo ‘drops his hotdog’, and el anda con maletas imaginarias ‘he’s walking with imaginary suitcases’. these comments ascribe effeminate gender characteristics to a male in order to convey the conversational message that he is homosexual. although these comments allude to the referent’s gender nonconformity, albeit in some roundabout way, they are understood in practice to mean simply, “he is gay”. that is to say, the actual conversational meaning these comments convey in practice is that of sexual, and not gender, non-normativity. this indicates that conventionalized heterosexist discourse in chile often relies on reference to gender non-normativity in order to imply non-normative sexual orientation. i will also examine the three phrases in the third category that position the referent in the passive sexual role: el muerde la almohada ‘he bites the pillow’, le gusta por el camino de tierra ‘he likes to take the dirt road’, and le gusta por detroit ‘he likes it by ‘detroit’’. these comments position the referent in the passive sexual role of the ‘recipient’, which is typically conceived as the female role. by verbally placing a man in the sexual role that dominant society has reserved for women, these comments derogate male homosexuality, conveying the message that gay males are inferior. all of these comments tend to be used frequently in everyday discourse, principally by gender-normative males. an aspect of identity in chilean society that is revealed by these verbal insults can be found in the syntactic composition that prototypically initiates such phrases. of the eight comments i examine in this paper, four begin with the passive-reflexive construction se le, which is an impersonal subject pronoun followed by an indirect personal pronoun. passive-reflexive sentences are commonly used in spanish as a way of de-emphasizing a person’s active role in a situation. for instance, a teacher who is accused by her students of forgetting to bring in their graded exams when she had promised to do so might defend herself by saying, “no me olvidé los exámenes, se me quedaron en la casa” (“i didn’t forget the exams, they stayed in the house on me”). stated this way, the teacher implies that it is not her fault that she does not have the students’ exams with her. i argue that this same type of construction, when used in a conventionalized heterosexist verbal insult, conveys the idea that the referent is a ‘victim’ of the unfortunate scenario in which the phrase involves him. for example, se le da vuelta el paraguas literally means ‘his umbrella is getting inverted on him’, and the referent is the victim of the negative event of his umbrella becoming broken. 3 3 balder: marginalization of alternative gender and sexual identities published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) likewise, in the phrase se le quema el arroz, which can be more literally translated as ‘his rice is getting burnt on him’, the referent is not actively burning his rice, but is rather the ‘victim’ of the unfortunate event of his rice getting burnt. i argue that, in conventionalized heterosexist comments, the use of passive-reflexive syntax is indicative that chileans conceptualize the ‘abnormality’ of alternative gender and sexual orientation as the outcome of something that went wrong. in chilean society, gender and sexual minorities are commonly described as victims of social or biological misfortunes that render them socially inadequate. conceptualized as victims by the dominant heteronormative culture, individuals with alternative gender and sexual orientation are in turn placed in a subordinate position within the social hierarchy. the resultant power differential between normative and non-normative individuals comprises the very essence of gender relations in society. cultural anthropologist roger lancaster, who maintains that the negotiation of social relations always involves power, states that the application of normative values to social relations results in, “not only an array of gendered bodies but also a world built around its definition of gender and its allotment of power” (1992:20). the conceptualization of gender and sexual minorities as essentially being rendered ‘woman-like’, which is reinforced by the use of heterosexist commentary in everyday discourse, denies non-normative males of power, social status, upward mobility, as well as the freedom to openly identify as gay. 3. linguistic mechanisms of verbal insults before returning to the sample comments on non-normativity, i would like to illustrate that these comments rely on combinations of a variety of linguistic mechanisms, such as metaphor, metonymy, and imagery. george lakoff, who has published widely on the subject, defines metaphor as “mappings from one domain to corresponding structures in another domain” (lakoff 1987:114). within this definition is the notion that speakers use their knowledge about one domain in order to understand and reason about another domain. a difference has been discovered between perceptual metaphors, which rely on surface or superficial similarities between two domains, and nonperceptual metaphors, in which the similarities between domains may be deeper, more occult, and less obvious (see vosniadou & ortony 1989; winner et al [1979]1993). winner et al provide an explanation of the latter type: “nonperceptual metaphors are based on relational similarities that cannot be apprehended by our senses. such metaphors are based on similarities between objects, situations, or events that are physically dissimilar but, often owing to parallel internal structures, function in a similar way” ([1979]1993:432). several of the heterosexist comments i collected rely on the mechanism of nonperceptual metaphor. of those that i discuss in this paper, the first four rely on the metaphor ‘homosexuality is an unfortunate situation’, whereby certain nonperceptual qualities of the source (a man in an awkward 4 4 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/3 doi: https://doi.org/10.25810/j1vk-ck75 marginalization of alternative gender and sexual identities social situation), are mapped onto the target (an effeminate homosexual man). such a metaphor attests that the way in which chileans perceive gender and sexual non-normativity is by equating it with their understanding of social dexterity, or more specifically, lack thereof. this metaphor shows that the chilean cultural understanding of homosexuals and effeminate men is that they are deficient and defective. john saeed defines metonymy as “identifying a referent by something associated with it” (1997:352). metonymy operates in that an x-like quality of the referent indicates that the referent is x. in the first three phrases i examine in this paper, the quality of social ineptness is identified in order to indicate that the referent himself is an inept member of society, i.e. a gay or effeminate man. it is important to note that, fundamentally, a vital component of being a ‘man’ in chile (by this i mean adequately fulfilling all of the social expectations of the male gender) includes the ability to be able to dominate certain situations. for example, men are expected to be able to financially support their families, perform sexually with women, solve problems, and excel at all of the other tasks demanded of them by society. inability to do these things depreciates a man’s masculinity, rendering him incompetent of fulfilling the expected masculine role in society. in the next two comments, the referent is metonymically identified as homosexual by his display of effeminate gesture, which is associated with homosexuality. the last three comments i examine in this paper are metonymic in that they rely on allusion to sexual passivity, which is a trait associated with male homosexuality, to identify the referent as a homosexual male. due in part to the creative nature of human speech, many of the comments i collected do not represent canonical examples of either metaphor or metonymy. because of this, it may be useful to apply the blend model, whereby both mechanisms contribute to the phrase’s composition and function. louis goossens (1995) has labeled this phenomenon of blending metaphor and metonymy in mental space as metaphtonomy. he asserts that, “although in principle metaphor and metonymy are distinct cognitive processes, it appears to be the case that the two are not mutually exclusive. they may be found in combination in actual natural language expressions” (goossens 1995:159). the heterosexist comments i collected are replete with multiple strata of semantic and pragmatic minutiae. goossens’ (1995) theory of metaphtonomy seems applicable to many of the expressions i came across. i examine possible applications of metaphor, metonymy, and metaphtonomy in the following analysis. 4. analysis of heterosexist comments: metalinguistic meaning and the dissonance between literal and intended meaning in this section, i expose the metalinguistic meaning of eight verbal insults, within the phrases in the first category, ‘allusion to gender non-normativity’, the link is drawn between the referent’s social or situational ineptness and his gender non-normativity. as gender non-normativity is equated with homosexuality in 5 5 balder: marginalization of alternative gender and sexual identities published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) dominant chilean society, these comments can effectively be used to label someone as homosexual. the first phrase, se le da vuelta el paraguas, ‘his umbrella gets inverted’, positions the referent in an inopportune situation – that of one’s umbrella being blown so that it flips outward, thus ceasing to shield its owner from the rain. the phrase’s nonperceptual ‘homosexuality is an unfortunate situation’ metaphor is easily recoverable. metonymically, this phrase uses the referent’s involvement in a non-normative and socially awkward situation to assign those very qualities to the referent. further, the referent’s awkwardness is mapped to his gender and sexuality, and thus the comment metaphtonymically conveys that the referent is an effeminate homosexual. the second phrase in this category, se le queda la pata atrás, which roughly means ‘his foot gets left behind’, is one of the most widely used heterosexist comments in everyday chilean discursive practices. in this phrase, the referent is again involved in an awkward situation, in which he is lame or walking in an ungainly manner due to a bum foot. the same train of metaphoric and metonymic reasoning with which i analyzed se le da vuelta el paraguas can be applied to this comment: by using the personal quality of social ineptness to indicate that the referent is an effeminate gay male, this phrase illustrates that social non-normativity in the general sense is used in order to conceptualize men with non-normative gender and sexuality. the next comment, se le quema el arroz (‘his rice is getting burnt’), also aligns with the ‘homosexuality is an unfortunate situation’ metaphor, in that it is a decidedly negative circumstance if the rice you are cooking gets scorched. participants in my study hypothesized that since cooking rice is typically considered a feminine activity, the phrase references a situation that is socially disadvantageous in the general sense, and also with respect to gender norms. the referent is essentially trying to engage in a feminine activity, which defies normative social gender expectations. this, plus the fact that he is unable to adequately succeed at cooking rice, indicates the referent’s awkwardness both situationally and with respect to ‘appropriate’ gender expression for males. furthermore, this comment plays on the feminine gender trait of overemotional reaction when small things go wrong. a stereotypical woman might become emotional or upset if her rice gets burnt, e.g. she might shriek ¡ai, se me quema el arroz! (‘oh! my rice is burning!’), instead of reacting with indifference as a typical man might (in the off-chance, that is, that he would even be in the kitchen in the first place). thus, when the comment is applied to a male referent, the implication is made that he, too, would react in a similar, overly emotional way to the type of miniscule misfortunes that women are popularly conceptualized as not being able to deal with. metonymically, the referent’s involvement in this disadvantageous situation signifies that the referent himself is somehow a disadvantaged member of society; and as ineptness or inadequacy are mapped to the traits of homosexuality and effeminacy, the resultant outcome is that the referent is indicated to be gay. 6 6 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/3 doi: https://doi.org/10.25810/j1vk-ck75 marginalization of alternative gender and sexual identities the first observation that can be made about se le cae el completo, which is in the category ‘reference to effeminate physical gestures’, is that it aligns with the ‘homosexuality is an unfortunate situation’ metaphor. however, cultural knowledge reveals the insight that this phrase also references an effeminate physical gesture. by relying on knowledge of cultural stereotypes, a mental image of the gesture referenced by this comment can be recalled. let me explain this gesture: imagine a man is holding a hotdog in what i will call the ‘resting position’ between taking bites, and then picture his hand and wrist bending away from him as the hotdog falls to the ground. this motion represents a stereotypically effeminate gesture, and its inclusion in this phrase contributes an element of humor above that provoked by the ‘homosexuality is an unfortunate situation’ metaphor. during my fieldwork in chile, it became apparent that humor is definitely a key element of chilean heterosexist discourse. in any case, the imagined gesture allows the phrase to function metonymically, in that the referent is identified as homosexual by something associated with homosexuality: display of effeminate gesture. although the other phrase in this category, el anda con maletas imaginarias, does not rely on the ‘homosexuality is an unfortunate situation’ metaphor, i feel that it takes the previously-mentioned humor component to an even greater height. this comment is metonymic in that again the referent is identified as homosexual by the effeminate gesture he demonstrates. however, the gesture referenced by this metonym is accessed by a different image. to grasp the image of the gesture invoked by this comment, one must visualize a man walking along, but instead of swinging his arms naturally in time with his stride, he holds them extended down the sides of his torso with his fists cocked upward as if each one was clutching the handle of an imaginary suitcase. according to chilean ideals about gender expression, this fashion of holding one’s hands and arms while walking is effeminate. and since use of effeminate gestures is associated with homosexuality, this comment, like the previous one, metonymically refers to a man as gay. the comments in the last category, ‘allusion to homosexual acts or homoerotic desire’, are principally metonymic in that they identify the referent as homosexual by calling to mind something that is associated with homosexuality – in this case, participation in a homosexual act or homoerotic desire. three of the phrases in this category specifically indicate that the referent is the passive participant in a homosexual act of anal sex. the first of these phrases, el muerde la almohada (‘he bites the pillow’), invokes the mental image of homosexual intercourse in which one man is being anally penetrated by another. the referent is depicted as the passive participant, and is positioned face down at the head of the bed where he can bite the pillow as he is being penetrated by the man on top of him. the other two phrases that touch upon the referent’s sexual passivity, le gusta por el camino de tierra and le gusta por detroit, metonymically draw upon the referent’s preference for being the recipient in anal sex acts in order to indicate that he is homosexual. specifically, they indicate that the referent prefers 7 7 balder: marginalization of alternative gender and sexual identities published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) to be anally penetrated by another man. each phrase indicates this preference in a slightly different way. in the first comment, el camino de tierra (‘the dirt road’) is a euphemism for anus or anal cavity. thus, the comment indicates that the referent likes sexual penetration by way of the anal cavity. the second phrase, le gusta por detroit (‘he likes it by ‘detroit’’), uses a play on words to allude to the referent’s preference for anal penetration. the city name ‘detroit’ is used as a euphemism for the phonetically similar word detrás, which means ‘behind’. thus, this comment alludes to the referent’s desire to be the passive participant in a homosexual act, which metonymically indicates that he is homosexual. all of these comments are non-literal figures of speech, in that their literal meaning clashes with their intended meaning: what is meant differs from what is said. although the phrases are observably complex, their metalinguistic meaning is oftentimes not fully comprehended by interlocutors who use or react to them. when used in discourse as verbal insults, hearers can immediately recognize that the utterance’s literal meaning is simply not plausible in the context at hand: in reality, no one’s rice is getting burnt, no one’s umbrella is inverted, no one is biting a pillow, no one is really walking down a dirt road, and even the imaginary suitcases are being imagined. sometimes, hearers already possess or are able to deduce the metalinguistic meaning that allows them to draw the connection between the literal and non-literal meaning of the phrase. other times, however, the hearers bypass the phrase’s metalinguistic meaning and only comprehend the intended meaning. by relying on communicative and cultural competence, interlocutors can employ their knowledge of chilean sociocultural beliefs and practices to navigate non-literal language metalinguistic meaning they possess no awareness of. 5. address terms in addition to this type of conventionalized verbal commentary, the equation of homosexual men with women is also commonly expressed in chile through the frequent use of address terms such as maricón, maraca, culiado, and hueco. these terms are loaded with metonymic reference to gender and/or sexual non-normativity, and play a vital role in the continuity of heterosexism’s prevalence in chilean society. the first term i would like to address is maricón, which is by far the most frequently used heterosexist address term. a diachronic approach shows that maricón is derived from the proper name maria. accordingly, this female origin qualifies the term to connotatively recall feminine gender attributes such as weakness and submissiveness. another parallel is that the term indicates submission not only in the social sense, but also in the sense of sexual passivity. in practice, this term is polysemous, and is primarily used to denote an effeminate and/or homosexual male, making it roughly synonymous with ‘effeminate sodomite’. i will refer to this primary meaning as (maricón 1). 8 8 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/3 doi: https://doi.org/10.25810/j1vk-ck75 marginalization of alternative gender and sexual identities additionally, this term can also be used secondarily in reference to a bad, wretched, or harmful person (maricón 2). the relationship between these two senses can be best explained by the chaining approach (see lakoff 1987), whereby maricón2 developed as a derivative of some of the characteristics of maricón 1. the characteristics, specifically, are those that give maricón 1 its negative connotation: maricón 1, which denotes an effeminate homosexual male, is connotatively negative in that homosexuals are considered bad, wrong, weak, deviant, and contemptible by dominant society. the chaining involved in the relationship of maricón 1 to maricón 2 can be conceptualized as such: maricón 1: effeminate passive homosexual male ↓ effeminacy, passivity, and homosexuality are bad, deviant, and wrong ↓ a maricón is a bad, deviant person who commits wrongful acts ↓ maricón 2: a bad, wretched, or harmful person the resultant polysemous outcome is that a term used for homosexuals can also be used for other people who are disreputable, odious wrongdoers (maricón 2), even if they have normative gender and sexual orientation. other such labels for homosexual males are maraca, culiado, and hueco. maraca (also sometimes pronounced marica) is a more antiquated form of maricón, and conveys a similar combination of negativity, effeminacy, and [homo]sexual passivity. the next term, culiado, in its most literal english translation, means “fucked”, or more specifically, anally penetrated. (culiado is a past participle of the verb culiar, ‘to anally penetrate’, which is derived from the noun culo, ‘anus’, or more colloquially, ‘ass’). the term culiado is metonymic, and draws upon passive participation in a homosexual act to indicate that the referent is homosexual. use of this term draws upon the most contemptuous aspect of male homosexuality, namely, sexual passivity. as such, it is loaded with negative connotation, and there are few instances where this term is not offensive. similarly, the term hueco also alludes to anal penetration. this term is polysemous: the primary denotation can be adjectival or nominal, and it means, respectively, ‘hollow’ or ‘something hollow’ (hueco1). the secondary meaning is a homosexual man (hueco2), and there is an underlying implication that he is the passive recipient in an anal sex act. hueco2 is derived from assigning the perceived qualities of something hollow or cavernous denotated by (hueco1) onto 9 9 balder: marginalization of alternative gender and sexual identities published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) a physical orifice that possesses those same qualities. in this sense, hueco2 is a synonym for ‘anus’ or ‘anal cavity’. as it denotes a homosexual male, hueco2 is synecdochic, whereby a body part is used in place of the whole entity – in this case, a man’s anal cavity is used in place of the man himself. the overall effect is that this term reduces the homosexual male to a vessel for penetration, and is thereby demeaning and derogatory. these four terms circulate prevalently in everyday discursive practices in chile, thereby keeping the popular conceptualization of homosexual males as effeminate and/or sexually passive intact within contemporary society. 6. conclusion: heteronormativity and dominant social values normative heterosexist discursive practices demote alternative sexual identities. as such, they provide verbal tools with which speakers can express their adherence to the normative gender and sexual expectations of dominant society. by derogating individuals who do not fall within the inventory of socially acceptable identities in chilean society, this type of language assigns legitimacy only to gender-normative heterosexuals, thereby denying gender and sexual minorities of social power. in this sense, active validation of socially normative sexual and gender values is synonymous with power in chilean society. in his article on heterosexual masculinity in college fraternities, scott kiesling exposes this negotiation of power in his definition of the discourse of heterosexuality, stating that, “heterosexual identities are not just displays of difference from women and gay men; they are also displays of power and dominance over women, gay men, and other straight men. a discourse of heterosexuality involves not only difference from women and gay men, but also the dominance over these groups” (2002:250). disguised as a creative variety of joking remarks, chilean heterosexist commentary reinforces the social structure in which genderand sexually-normative men maintain their dominant role, causing gender and sexual minorities to be marginalized. the fact verbal insults are most often used in a joking manner, as normative men poke fun at each other or at other non-normative males, is in and of itself problematic: because chileans who use these comments consider the practice to be ‘just joking around’ or hueveando, they tend to deny all accusations that they are homophobic, or that they discriminate against gays. as one of my research informants stated, los chilenos son homofóbicos sin darse cuenta – ‘chileans are homophobic without even realizing it’. while it is true that the attitudes towards gays in chile do not extend to outright legal repression or violent intolerance, they do nevertheless present a very detectable underlying current of discomfort for sexual minorities. relatively few chileans disclose their alternative sexual orientation to their family, and even fewer have come forth to publicly identify as homosexual. sexual minorities who cannot openly identify as such are called tapado, which means ‘capped’, ‘concealed’, or ‘covered up’. this 10 10 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/3 doi: https://doi.org/10.25810/j1vk-ck75 marginalization of alternative gender and sexual identities 11 metaphoric label is indicative of the social restraints that cause sexual minorities to feel constricted within the limited confines of a society governed by heterosexist beliefs and practices. due to its prevalence in everyday discourse, heterosexist language plays a vital role in the general knowledge within chilean society of the cultural norms and expectations surrounding gender and sexual identity. this assertion aligns with the opinions of sociolinguists mary bucholtz and kira hall, who stated that, “language is a primary vehicle by which cultural ideologies circulate, it is a central site of social practice, and it is a crucial means for producing sociocultural identities” (bucholtz & hall 2004:512). to this respect, i argue that by positioning non-normative individuals as the butt of jokes that derive their humor from synonymizing homosexual males with women, chilean vernacular represents one of the most powerful forces that cause the reproduction of heterosexist beliefs in society, resulting in the delegitimization – and consequent marginalization – of alternative identities. references bucholtz, mary, and kira hall (2004). theorizing identity in language and sexuality research. language and society. bucholtz, mary, and kira hall (forthcoming). identity and interaction: a sociocultural linguistic approach. discourse studies. goossens, louis (1995). metaphtonomy: the interaction of metaphor and metonymy in figurative expressions for linguistic action. in goossens, louis, et al, by word of mouth: metaphor, metonymy, and linguistic action in a cognitive perspective, 159-174. philadelphia: john benjamins publishing company. kiesling, scott f. (2002). playing the straight man: displaying and maintaining male heterosexuality in discourse. in kathryn campbell-kibler et al., eds., language and sexuality: contesting meaning in theory and practice, 249266. palo alto, ca: csli publications. lakoff, george (1987). women, fire, and dangerous things: what categories reveal about the mind. chicago: university of chicago press. lancaster, roger n. (1992). life is hard: machismo, danger, and the intimacy of power in nicaragua. berkeley: university of california press. marrones, nila (2001). de colores: lesbianas y gays latinos: historias de fuerza, familia y amor. www.unlearninghomophobia.com saeed, john i. (1997). semantics. oxford: blackwell publishing. vosniadou, stella and andrew ortony (1989). similarity and analogical reasoning. new york: cambridge university press. winner, ellen, and howard gardner. ([1979]1993). metaphor and irony: two levels of understanding. in andrew ortony (ed.), metaphor and thought, 425-445. philadelphia: cambridge university press. 11 balder: marginalization of alternative gender and sexual identities published by cu scholar, 2005 http://www.unlearninghomophobia.com/ colorado research in linguistics 6-2005 marginalization of alternative gender and sexual identities: the role of normative discursive practices in chilean society sara rose balder recommended citation cril an act-r model of sentence sorting with argument structure constructions colorado research in linguistics. june 2005. vol. 18, issue 1. boulder: university of colorado. © 2005 by anna m. fowles-winkler and laura michaelis. an act-r model of sentence sorting with argument structure constructions∗ anna m. fowles-winkler and dr. laura michaelis university of colorado based on the results of a sorting task involving verbs and grammatical patterns, bencini & goldberg (2000) argue that “argument structure constructions are directly associated with sentence meaning.” we explore this hypothesis by attempting to replicate their results using a nonhuman categorizer: a cognitive model based on act-r (anderson & lebiere 1998). the model replicated the sentence-sorting behaviors of bencini & goldberg’s subjects, but did so using formal cues alone. this outcome suggests that the subjects in the bencini & goldberg study were not necessarily attending to constructional meaning, and lends support to bock’s (1986) conclusions regarding syntactic priming: subjects’ similarity judgments are as likely to be based on syntactic form alone as they are to involve syntax-semantic mapping. 1. introduction this paper investigates whether argument structure patterns play a role in sentence interpretation. bencini and goldberg argue, based on two experiments described in their paper, that “argument structure constructions are directly associated with sentence meaning” (bencini & goldberg 2000). this constructional approach differs from other approaches, specifically, lexical and multiple-sense approaches. a lexical approach posits that syntactic and semantic information are encoded entirely on the verb. a multiple-sense approach views different verb constructions as different representations for that verb. bencini and goldberg conducted experiments in which subjects sorted sixteen sentences by meaning. the sixteen sentences were comprised of four different verbs, and four different constructions. based on previous sorting experiments (regehr & brooks 1995), it would be expected that subjects would sort on a single dimension—by verb. however, bencini and goldberg discovered that subjects took both verb and construction into account when considering sentence meaning. we explore this hypothesis by attempting to replicate bencini and goldberg’s results using a nonhuman categorizer: a cognitive model based on ∗ ms. fowles-winkler would like to thank brad best of micro analysis and design for discussing act-r modeling and dr. adele goldberg for kindly mailing a hardcopy of her paper “relationships between verbs and constructions.” 1 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 2 act-r (adaptive control of thought-rational) (anderson & lebiere 1998). experiments were conducted with the act-r model following bencini and goldberg’s experiment method. the first experiment resulted in sorts that indicate that the model was sorting primarily by verb. the second experiment resulted in a divided strategy, one that was neither close to a verb sort, nor close to a constructional sort. this paper is structured as follows: in section 2, the bencini and goldberg experiment is described, including a comparison of the constructional model to the lexical and multiple-sense approaches. section 3 details sorting strategies often employed by subjects: single-dimension and family resemblance sorts. section 4 provides a basic description of the act-r cognitive modeling framework. in section 5, the act-r model created for this experiment is described. section 6 presents the experiment results. finally, section 7 discusses the results of the act-r model experiments in comparison to the findings of bencini and goldberg, and presents future issues for consideration. 2. the bencini and goldberg experiment bencini and goldberg conducted experiments to “test whether argument structure constructions play a role in determining sentence meaning” (b&g 2000: 643). a verb’s argument structure describes the number and types of participants defined for that verb. for example, in sentence (1) below, the verb ‘sing’ has one argument, mary, an actor. (1)mary sings. an argument structure construction is a verb-level grammatical pattern that links semantic roles to grammatical roles (goldberg 1995, michaelis & ruppenhofer 2001). event-structure meanings classically viewed as the output of lexical rules, e.g., the ditransitive or ‘double object’ pattern, are instead viewed as construction meanings. the meaning of a sentence arises from the integration of the verb’s meaning and the construction’s meaning as, e,g., she sliced the tomatoes into the salad denotes the means by which the agent of the sentence effected the caused motion event denoted by the construction (bencini & goldberg: 642). this construction-based model used by bencini and goldberg is based on construction grammar. construction grammar is a theory that looks at all components of language—core and non-core language uses, and does not have a strict division between the lexicon and syntax, or between semantics and pragmatics (goldberg 1995: 7). this theory is committed to treating “all types of expressions as equally central to capturing grammatical patterning (i.e. without assuming that certain forms are more ‘basic’ than others) and in viewing all dimensions of language (syntax, semantics, pragmatics, discourse, morphology, 2 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/1 doi: https://doi.org/10.25810/spvf-m355 an act-r model of sentence sorting with argument structure constructions 3 phonology, prosody) as equal contributors to shaping linguistic expressions” (“construction grammar”). an alternative approach to the construction model is the lexical projection model. this approach theorizes that the verb is the primary source of sentence comprehension—it holds syntactic and semantic information. different uses of the same verb are derived by using lexical rules or transformations. for example, the ditransitive/prepositional alternation is produced by a lexical rule that transforms semantic structure. pinker suggests that this rule takes an input verb “with the semantics ‘x causes y to go to z’ and produces the semantic structure ‘x causes z to have y’” (goldberg 1995: 8). one problem with this approach is that there are ditransitive expressions that do not have a corresponding prepositional expression, and vice versa: (2)a. jane refused fred a kiss. (goldberg 1992) b. *jane refused a kiss to fred. (3)a. i said my prayers to my mother. b. *i said my mother my prayers. another problem with the lexical approach is that some verbs do not entail that z has y as illustrated by the examples in (4): (4)a. rose flipped the pancake to cleo, but it landed on the floor. b. ron threw the ball to cal, but he didn’t get it. in both examples, the entailment that z has y is defeased: cleo does not have the pancake, nor does cal have the ball. another approach is a theory that suggests different verb senses for a single verb. bencini & goldberg refer to this theory as the multiple-sense approach. rappaport hovav and levin describe this approach by explaining “variations in a verb’s meaning are because of a verb’s basic semantic classification being expanded” (rappaport hovav and levin 1998: 104). a verb has multiple lexical semantic templates—or event structure templates (rappaport hovav and levin 1998: 107). these multiple templates account for multiple verb meanings. for example, these are a few of the variations of the verb ‘sing’: (5)a. darla sang. b. darla sang the song. c. darla sang the song to charlie. d. darla sang the song across the ocean. e. darla sang the song loudly. these different uses, or senses, of the verb are captured by different event structure templates. 3 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 4 one benefit to using the constructional approach is multiple (usual and unusual) senses of verbs are avoided. consider sentence (6): (6)sam sneezed the napkin off the table. (goldberg 1995: 29) the verb ‘sneeze’ is typically an intransitive verb, however, based on the above example, a lexical approach would have to suggest that ‘sneeze’ takes three arguments. a constructional approach, on the other hand, would argue that the construction itself contributes the change in meaning, not the individual verb. another benefit to associating sentence meaning to both the verb and the construction is seen with a sentence that contradicts the meaning of the construction, as in (7): (7)pat ignored chris. in (7), the transitive construction is x acts on y. however, the meaning of the sentence is clearly that pat (x) has nothing to do with chris (y). the verb ‘ignore’ negates the transitive construction. according to goldberg (1997), “the meaning of the verb is integrated with the meaning of the construction, resulting in entailments that neither the verb or the construction have independently.” without the relationship between verb and construction, the meaning of the sentence would be based solely on the verb, disregarding the meaning provided by the construction. in the bencini and goldberg study, subjects were asked to sort 16 sentences into four piles by meaning; these sentences combined four different verbs (throw, get, slice, take) with four different constructions (transitive, ditransitive, caused motion, resultative). table 1 lists the experiment stimuli. our experiment uses the same stimuli. in bencini and goldberg’s first experiment, 17 participants were tested as a group and were asked to write a paraphrase for each sentence. they were then asked to sort the sentences into four piles based on the sentence meaning. the participants were also told that the same words in sentences can have different meanings, as seen in the examples “kick the bucket” versus “kick the dog” (bencini & goldberg 2000). of the 17 participants, 7 sorted by construction alone, and 10 used a mixed sort strategy. none of the participants sorted only by verb. because the experiment instructions could have influenced the participants’ avoidance of verb only sorts, bencini & goldberg conducted a second experiment. the second experiment involved the same number of participants. the procedure was modified slightly as follows: participants were tested individually, there was no mention of the same words having different meanings, and participants were asked to explain their grouping strategies after they completed the task. in this experiment, 7 sorted by verb, 6 by construction, and 4 used a mixed sort strategy. bencini and goldberg analyzed the explanations given by 4 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/1 doi: https://doi.org/10.25810/spvf-m355 an act-r model of sentence sorting with argument structure constructions 5 the participants to determine if they were attending to sentence meaning or surface cues. one example of an explanation given for a ditransitive sentence is: “here one person is doing something for another person” (bencini & goldberg 2000). based on that explanation, the participant could have been attending to the overall sentence meaning. however, all of the ditransitive stimuli involved two people, so participants could have been paying attention to surface cues only. table 1: experiment stimuli verb transitive ditransitive caused motion resultative throw (8) anita threw the hammer. (12) chris threw linda the pencil. (16) pat threw the keys onto the roof. (20) lyn threw the box apart. get (9) michelle got the book. (13) beth got liz an invitation (17) laura got the ball into the net. (21) dana got the mattress inflated. slice (10) barbara sliced the bread. (14) jennifer sliced terry an apple. (18) meg sliced the ham onto the plate. (22) nancy sliced the tire open. take (11) audrey took the watch. (15) paula took sue a message. (19) kim took the rose into the house. (23) rachel took the wall down. although previous results (e.g., regehr & brooks 1995) lead to the prediction that subjects would perform unidimensional sorts (i.e., by verb), the results of the bencini & goldberg experiment suggested that subjects took both verb and construction semantics into account. there is, however, an alternate interpretation of their results: subjects were not attending to event-structure semantics, but were instead performing pattern matching. in order to determine which of these two construals is correct, an act-r model was created that uses random selection and pattern matching to sort sentences. 3. typical sorting strategies while this experiment primarily examines whether subjects attend to verb and construction semantics or surface cues, a second but related question is how do subjects decide to categorize sentences: what four piles are considered, and which strategies are selected? 5 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 6 when subjects are asked to categorize a group of objects with no feedback or instructions, the task is called category construction (also “free sorting” or “free classification”) (milton & wills 2004: 407). the task described in this paper is a modification of category construction, as the act-r model is required to create four piles. in determining the heuristics to encode in the model, literature about category construction was reviewed. there are two basic methods employed by subjects to sort objects into categories: unidimensional or family resemblance. a unidimensional, or single dimension, sort occurs when a subject picks one attribute or feature about the stimuli, and creates categories based on that attribute. a family resemblance sort occurs when the subject identifies a number of dimensional values that the objects have in common. this is “a sort pattern in which one prototype and all of its derived one-aways are placed in one category” (regehr & brooks 1995: 349). in one set of experiments, it was found that the experiment procedure influenced how the subjects sorted sentences. a match-to-standards procedure reduced the use of unidimensional sort, and produced a family resemblance sort. this procedure displayed objects one pair at a time to subjects, and prevented subjects from viewing the entire array of objects at one time. regehr and brooks speculate that the “tendency toward family resemblance sorting seems to be a function of the fact that participants were focused on pairwise comparisons of objects” (1995: 355). in their experiments, they discovered that a full stimulus array discourages family resemblance sorting. in other words, by showing subjects all of the objects to sort, the subjects would not choose to sort by family resemblance, but by a single dimension. however, milton and wills (2004) found that a match-to-standards procedure does not always produce a family resemblance sort. in their experiments, they determined that both procedure and experiment stimuli impact the sorting method chosen by subjects. these findings indicate that it is not necessarily clear which sorting method the procedure and stimuli chosen for this experiment would produce: a unidimensional (verb-based) or family resemblance (construction-based) sort. we conducted a small, informal investigation to determine heuristics people employ when presented with the sentence-sorting task. ten subjects were asked to sort the sentences into four piles based on sentence meaning. no other instruction was provided. three subjects sorted by verb. two subjects sorted by construction. the other five subjects used mixed sorts. the two constructions that subjects seemed to easily identify are the ditransitive and the resultative constructions. one subject mentioned that the ditransitive sentences all involved two people. these anecdotal results are confirmed by bencini and goldberg’s results—the ditransitive is the easiest construction to identify. also, it appeared that the caused motion and resultative constructions were occasionally grouped together, indicating that they were understood by subjects to be similar. bencini and goldberg also confirm that these two constructions are closely related. 6 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/1 doi: https://doi.org/10.25810/spvf-m355 an act-r model of sentence sorting with argument structure constructions 7 based on anecdotal evidence and bencini and goldberg’s results as described in the above paragraph, two mixed sorting strategies were created for the act-r model. these strategies are discussed in section 5 below. 4. a brief description of act-r act-r is a cognitive modeling framework—it is a theory of human cognition. the current version of act-r (act-r 5.0) is based on over 20 years of work, and many previous versions of the theory; therefore, a very brief description of act-r is provided here. act-r has three primary components: modules, buffers, and a pattern matcher. there are two kinds of modules: perceptual-motor and memory. the perceptual-motor modules handle how the model interacts with the world via visual, auditory, or motor mechanisms. the memory module consists of declarative memory and procedural memory. declarative memory is knowledge that we can articulate to others, represented by chunks. the following is a chunk used for this experiment: (chunk-type meaning word type) the statement declares a chunk called meaning with two slots, word and type. a slot is an attribute of the chunk, and the values of the slots define the chunks. the meaning chunk was used to encode the experiment stimuli, with each word encoded. for example: (hammer isa meaning word “hammer” type object) procedural memory represents what we know to do with the declarative memory, and is represented by productions. a production consists of a series of buffer tests, then of a series of buffer transformations. the buffer tests are one of the ways act-r determines how to choose a production. the following is a production used for this experiment: (p process-first-noun =goal> isa comprehend-sentence agent nil action nil object1 nil object2 nil word =word state read =retrieval> isa meaning word =word 7 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 8 ==> =goal> agent =retrieval word nil state find) the production is called process-first-noun, and can be translated as follows: if the goal is to comprehend the sentence the model is reading and agent, action, object1, object2 are nil and word is not nil and, the retrieved chunk has the same value as word then change the goal by setting agent to the retrieved word read the next word (set the state to find) act-r uses a number of buffers that collectively represent the current state of the model. the goal buffer represents where the model is in completing the current task. the retrieval buffer, shown in the above production, contains chunks that have been retrieved from declarative memory. there are additional buffers for storing information obtained from visual and auditory channels, but since these buffers do not directly apply to this experiment, they are not described here. based on the current state of the buffers, the pattern matcher determines which productions may execute, thus causing the buffer contents to potentially change. for a more detailed description of the act-r framework, refer to anderson and lebiere (1998). 8 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/1 doi: https://doi.org/10.25810/spvf-m355 an act-r model of sentence sorting with argument structure constructions 9 5. the sentence sorting act-r model the act-r model for this experiment was created using act-r 5.01. conceptually, the model represents a human performing the sorting experiment. the model starts with the sentence words encoded in declarative memory. these are the words that match the agent, verb, object, and oblique portions of the experiment stimuli listed in table 1. the words “the,” “a,” and “an” are skipped because they are not necessary for this experiment. the model runs in two phases. the first phase reads all sixteen sentences, creating chunks for each sentence and storing those chunks in memory. the act-r model uses a very simple parsing mechanism to construct sentence chunks. a sentence chunk has the following slots: • agent: the first word in the sentence, corresponding to the subject; • action: the second word in the sentence, corresponding to the verb; • object1: the third word in the sentence (ignoring a, an, and the); • object1-type: the type of the object is encoded in declarative memory to indicate animate objects; • object2: the fourth word in the sentence (the second object); • preposition: the preposition in the sentence; • category: the selected pile to place the sentence; • sort-strategy: the sort strategy chosen by the model (verb, construction, or mixed) • purpose: study or categorize; • word: the current word being read by the model; • state: the current state of the model, used to direct production execution. because some slots are empty depending on the sentence, productions look for either empty slots, or a combination of empty and full slots to sort the sentences. this strategy is discussed below. after all sixteen sentences are read, the word “categorize” is shown to the model. at this point, the model randomly picks one of the three sorting strategies (verb, construction, or mixed). the three sorting strategies are represented by three individual productions; each has a random chance of being selected. the sorting strategy is stored on each sentence chunk, and the second phrase starts. the model reads the sixteen sentences again, and sorts each sentence into one of four piles using the chosen sorting strategy. 1 act-r 5.0 is freely-available software from carnegie mellon university, http://act-r.psy.cmu.edu/ 9 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 10 the verb sorting strategy matches the contents of the action slot to a verb. there are four productions that comprise the verb sort. the verbs are sorted into four piles as follows: • pile 1: throw • pile 2: get • pile 3: slice • pile 4: took. the construction sorting strategy consists of four productions for each construction. each production sorts the sentences based on surface cues. the transitive production matches sentence chunks with values in the agent, action, and object1 slots, and no values in the object2 and preposition slots. this template is agent action object1. the ditransitive production matches sentence chunks with values in the agent, action, object1, and object2 slots. the object1 slot must be a proper noun, which is coded by a meaning type of ‘animate’. the caused motion production matches sentence chunks with values in the agent, action, object1, object2, and preposition slots. additionally, the object1 is not a proper noun. finally, the resultative construction production is similar to the ditransitive production except that the object1 is not a proper noun. the construction sorting strategy sorts the sentences as follows: • pile 1: transitive • pile 2: ditransitive • pile 3: caused motion • pile 4: resultative. two different mixed sorting strategies were implemented. the first experiment’s mixed sorting strategy sorts the sentences as follows: • pile 1: threw or get verbs • pile 2: slice verbs • pile 3: take verbs • pile 4: resultative construction there are five constructions for this sorting strategy, all based on the previously discussed productions. the difference with the mixed strategy is that a combination of verb and construction productions is used, and it is possible for multiple productions to match certain sentences. when multiple productions match, a production is randomly selected. for example, sentence (23) rachel took the wall down would match both the take verb production, and the resultative construction. the model would randomly pick the production (i.e., the pile to put the sentence in), so some runs might have sentence (23) in pile 1, and others might have the sentence in pile 4. the second experiment uses a mixed strategy that sorts the sentences as follows: • pile 1: get or take verbs • pile 2: ditransitive construction 10 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/1 doi: https://doi.org/10.25810/spvf-m355 an act-r model of sentence sorting with argument structure constructions 11 • pile 3: throw or slice verbs • pile 4: caused motion or resultative constructions. this strategy uses seven productions, all of which are based on the previously discussed verb and construction productions. this strategy exhibits more randomness than the first strategy because it is possible for more sentences to match multiple productions. 6. experiment results 6.1. experiment 1 method participants. an act-r model was created and run 50 times to simulate 50 participants. stimuli. the sixteen sentences shown in table 1. procedure. for each run, the sentences were displayed sequentially, with the model reading each sentence. this is a change from the bencini and goldberg procedure in two ways. first, in the bencini and goldberg experiment, subjects were shown all sentences at the same time. this method probably influences how subjects create categories and determine sorting methods for the sentences. to simplify the act-r model, it was decided to show the sentences sequentially, and to have the sorting strategy chosen randomly. second, bencini and goldberg had their subjects write a short paraphrase of each sentence to ensure the sentences were processed, and possibly to indicate comprehension. this task was omitted from this experiment because this experiment focuses on producing similar sorting results with little or no comprehension. after reading all sixteen sentences, the word “categorize” is displayed to indicate to the model that the sorting should commence. at this point, a sorting strategy is randomly selected by the model. random strategy selection was chosen to try to simulate real world experiments where subjects would not all necessarily choose the same sorting strategy. next, the same sixteen sentences were displayed sequentially again. after the model read each sentence, the model sorted the sentence into one of four piles using one of three sorting strategies: verb, construction, or mixed. as mentioned above, the verb sorting strategy sorts entirely by verb, the construction sorting strategy sorts entirely by construction, and the mixed sorting strategy classifies sentences based on verb and construction. results of the 50 runs, 15 (30%) were verb sorts, 17 (34%) were construction sorts, and 18 (36%) were mixed sorts. these results are expected because the 11 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 12 model should have been selecting the sort strategy randomly. to analyze the mixed sorts, the verb and construction deviation scores used by bencini and goldberg were applied. the deviation score for a verb sort (vdev) is calculated by counting the number of changes that would be made to make the sort entirely verb-based. the maximum vdev score is 12. the deviation score for a construction sort (cdev) is calculated by counting the number of changes that would be made to make the sort entirely construction-based. the maximum cdev score is also 12. therefore, a verb sort has a vdev score of 0, and a cdev score of 12, while a construction sort has a vdev score of 12, and a cdev score of 0. there was a single run that did not sort the sentences into four piles (the sentences were sorted into three piles). this run was omitted from the data results because there was sufficient data without it. in examining the runs that produced mixed sorts, the vdev values are overall closer to 0 than the cdev values. the cdev values are closer to 12, indicating that the mixed sorting strategy in the model is most likely biased towards a verb-based sort. the mean vdev for the mixed sorts, or average number of changes required for the sort to be entirely verb-based, is 4.7 (σ = 0.9) the mean cdev for the mixed sorts, or average number of changes required for the sort to be entirely construction-based, is 9.7 (σ = 0.8). 6.2. experiment 2 method participants. the act-r model used for experiment 1 was modified for experiment 2, and run 50 times to simulate 50 participants. stimuli. the stimuli were the same as for experiment 1. procedure. the procedure for the experiment was the same as for experiment 1. however, after examining the results from experiment 1, the mixed sort strategy used by the act-r model was modified with the intent of producing lower cdev values, i.e., biasing the sort towards a construction sort. the ditransitive construction was added to the mixed sort strategy. this construction was chosen because, according to bencini and goldberg, this is the easiest construction to identify. additionally, the caused motion construction was added because this construction is closely related to the resultative construction (bencini & goldberg, 2000, p. 646). the modified mixed sort strategy is: pile 1: get or take verbs, pile 2: ditransitive construction, pile 3: throw or slice verbs, and pile 4: caused motion or resultative constructions. productions were added to the model to perform the sorting actions. as in experiment 1, multiple productions can match a particular sentence, so the model randomly picks one to fire. for example, sentence (16) pat threw the keys onto the roof could be placed in pile 3 or pile 4. the model randomly picks which pile to place the sentence. also, as in experiment 1, the model picks the sort strategy randomly from verb-based, construction-based, and mixed. results 12 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/1 doi: https://doi.org/10.25810/spvf-m355 an act-r model of sentence sorting with argument structure constructions 13 of the 50 runs, 11 (22%) were entirely verb-based sorts, 19 (38%) were construction-based sorts, and 20 (40%) were mixed sorts. these results are not surprising as the model picks the sorting strategy at random. the same verb and construction deviation scores discussed above were calculated for the mixed sorts. the verb and construction deviation scores for the mixed sorts were both fairly high, indicating that the sorting strategy was too divisive. in other words, it was not possible for the model to sort close enough to either a verb-based or construction-based strategy. the mean vdev is 8.1 (σ = 1.1), and the mean cdev is 7.3 (σ = 1.2). 7. discussion bencini and goldberg’s first experiment found that subjects were more influenced by constructions than verbs. the vdev value for their experiment is 9.8, and the cdev value is 3.2. their second experiment produced a vdev value of 5.5 and a cdev value of 5.7, indicating that sorts were divided between verb and construction strategies. in comparison, the act-r model used for experiment 1 was more influenced by verbs than constructions (vdev 4.705 and cdev 9.764). after modifying the mixed sort strategy, the model used in experiment 2 was divided between verb and construction sorts (vdev 8.105 and cdev 7.315). the high mean vdev and cdev values produced in experiment 2 suggest that further modifications of the model are required. these numbers potentially indicate that it is possible to create a simulation of a subject sorting experiment, and, at least in the case of experiment 1, produce results that seem plausible. the model randomly chose one of three sorting strategies: entirely by verb, entirely by construction, or mixed, by verb and construction. when the model sorts by verb, it looks only at the verb in each sentence, and groups sentences with the same verb into one pile. construction sorts use templates with checks for proper nouns and prepositions. the mixed strategy uses a combination of verb and construction sorts. because the mixed strategy used in experiment 1 resulted in more verb sorts, it was modified for experiment 2 to bias the strategy towards construction sorts. currently, the act-r model conflates category construction and sentence sorting in phase 2. the model could be modified to separate those tasks by constructing the four categories during the first phase of sentence parsing, and using those categories for sorting during the second phase. this modification would most likely use the stored sentence chunks, as the model would need to retrieve previously seen sentences from memory to revise the categories. these changes would seem to better parallel how a human subject would approach the sorting task. while the results presented in this paper are preliminary, they do indicate that there is much more to explore in this area. first, it would be worthwhile to compute the verb and construction spread to compare values for this experiment 13 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 14 with those produced by bencini and goldberg. second, by changing the experiment procedure to a match-to-standards procedure (one in which subjects only see two sentences at a time) it would be possible to compare an act-r model results with human subjects. finally, another consideration for future experimentation is to modify the stimuli used by adding additional sentences to have a total number that is not evenly divisible by 4. by using an even number of verbs and an even number of constructions, the stimuli might encourage subjects to sort along either verb or construction. in sum, the model replicated the sentence-sorting behaviors of b&g’s subjects, but did so using formal cues alone. this outcome suggests that the subjects in the b&g study were not necessarily attending to constructional meaning, and lends support to bock’s (1986) conclusions regarding syntactic priming: subjects’ similarity judgments are as likely to be based on syntactic form alone as they are to involve syntax-semantic mapping. 14 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/1 doi: https://doi.org/10.25810/spvf-m355 an act-r model of sentence sorting with argument structure constructions 15 references anderson, j. r., & lebiere, c. (1998). the atomic components of thought. mahwah, new jersey: lawrence erlbaum associates. bencini, g. m., & goldberg, a. e. (2000). the contribution of argument structure constructions to sentence meaning. journal of memory and language, 43, 640-651. bock, j. k. (1986). syntactic persistence in language production. cognitive psychology, 18, 355-387. construction grammar. (n.d.). retrieved april 17, 2005, from http://www.constructiongrammar.org goldberg, a. e. (1992). the inherent semantics of argument structure: the case of the english ditransitive construction. cognitive linguistics, 3(1), 37-74. goldberg, a. e. (1995). constructions: a construction grammar approach to argument structure. chicago: the university of chicago press. goldberg, a. e. (1997). relationships between verb and construction. in lexicon and grammar (pp. 383-398). john benjamins. michaelis, l. a., & ruppenhofer j. (2001). beyond alternations: a constructional model of the german applicative pattern. stanford: csli publications. milton, f., & wills, a. j. (2004). the influence of stimulus properties on category construction. journal of experimental psychology: learning, memory, and cognition, 30(2), 407-415. rappaport hovav, m., & levin, b. (1998). building verb meanings. in the projection of arguments: lexical and compositional factors (pp. 97-134). cambridge university press. regehr, g., & brooks, l. r. (1995). category organization in free classification: the organizing effect of an array of stimuli. journal of experimental psychology: learning, memory, and cognition, 21(2), 347-363. 15 fowles-winkler and michaelis: an act-r model of sentence sorting with argument structure constructions? published by cu scholar, 2005 colorado research in linguistics 6-2005 an act-r model of sentence sorting with argument structure constructions anna m. fowles-winkler laura michaelis recommended citation microsoft word fowles-winkler.anna.comments revised.doc between authority and authenticity: english use in spanish-language commercials in the united states colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by nikki steeby. between authority and authenticity: english use in spanish-language commercials in the united states nicole steeby university of colorado at boulder the use of english in foreign-language advertising abroad has been explored in depth, as has the role english plays in globalization. however, there remains a paucity of research concerning english use in foreign-language advertising within english-speaking countries. this analysis of miami-based spanish commercials explores the roles that identity and citizenship play in profitmotivated code-switching, in addition to questioning the oft-held assumption that english use carries a singular social meaning regardless of context. beginning with a look at previous advertising studies, namely those of ingrid piller and tej bhatia, this study examines how english language use within the united states both contrasts and coincides with its use in german and indian advertising. next, pierre bourdieu’s concept of a linguistic marketplace is introduced to analyze class relations as they relate to code-switching. the study then draws from mikhail bakhtin to explain how advertising employs heteroglossia and double-voicing to sell products and propagate existing ideologies of language value and class dominance. this is accomplished primarily by means of profit-motivated code-switching between english and spanish phonologies, which lend either authority or authenticity to products. lastly, a discussion of bonnie urciuoli and robin lakoff aids in understanding the concepts of the public and the private spheres as they relate to linguistic hegemony and marketing tactics. 1. introduction the use of english in advertising is a global phenomenon. as ingrid piller points out, “english is the most frequently used language in advertising messages in non-englishspeaking countries (besides the local language)” (2003: 175). studies of commercialized english have shown that the language is often used to index stereotypes of western culture or, more commonly, to index a generalized conception of modernity and progress. piller, in her studies of german advertising, has found that english is employed to target a cosmopolitan, upper-class market. she writes, “[t]he implied reader of englishgerman ads is ‘not a national citizen, but a transnational consumer’” (2003: 176). also, takashi, (as discussed in piller, 2003) has found that english in japan, “… does not index americanization, or even westernization, but rather indexes a modern, sophisticated, and cosmopolitan identity for the products” (175). however, tej bhatia, (in grant and short, 2002) in his research on rural india, has noticed a new trend precipitated by a backlash against english in transnational advertising. because many peoples fear that their indigenous cultures will be eradicated by globalization, a recent marketing tactic is what bhatia terms glocalization. the term refers to “the integration of localization and globalization” (70), and especially the manner in which international products are adapted to local markets. the above findings describe, and seek to understand, the prevalence of english in foreign advertising and globalization in general. however, a question that remains largely unexplored—in these and other studies—is how english is used in foreign1 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 language commercials in english-speaking countries, where english already enjoys a dominant position over other languages. reasonably, one might posit that english use would be even more prevalent in such a context. in the united states, for example, english is both the official and most widely spoken language, which legitimizes it on a legal as well as practical basis. nonetheless, my collection of 102 for-profit spanish television advertisements, which was amassed over a two week period in november, 2003, and broadcast from the miami station univision, does not conform to expectations. as i will demonstrate in this paper, english in spanish-language commercials in the u.s. does not always, or even often, enjoy the expected linguistic hegemony. perhaps one important reason for this is that english use in the u.s. is often complicated by questions of identity, including citizenship. in the three studies mentioned above, it is improbable that english use would have any bearing on the legal status of german, japanese, or indian nationals living in their respective countries. english use among u.s. hispanics, however, may be used to index citizenship or even ‘whiteness.’ this, in turn, can ultimately raise the question of allegiance to hispanic and national identities. therefore, many u.s. advertising companies shy away from english use to avoid alienating their target consumers and to gain a more widespread acceptance for their products. a second important difference in u.s. advertising is that, unlike the studies of germany or japan discussed above, u.s. spanish commercials specifically target a lessaffluent consumer base. the audience is largely lower to middle-class hispanic immigrants, many of whom are undocumented, monolingual spanish speakers. preliminary evidence for class is substantiated by the lack of luxury goods in commercials. in her study, piller found that many commercials directed at the german cosmopolitan elite consisted of ads for computers, suits, expensive jewelry and luxury automobiles (2001). however, not one commercial depicting any of these goods is to be found in my collection of spanish commercials. the two categories that come closest are advertisements for technology and automobiles. with respect to technology, internet commercials are becoming increasingly common, although there are still no ads for the computers themselves. automobile advertisements, in turn, are almost exclusively for pick-up trucks—which are not considered luxury goods from a dominant class perspective. the next demographic indicator relates to the hispanic audience itself, which is often stereotypically addressed and depicted in many of the advertisements. the law offices of frickey, for example, specifically target the working-class. they begin one ad by zooming in on a construction site that depicts hispanic workers. the ad addresses these men saying: you are a worker, proud, reliable. you work hard to support your family. (usted es un obrero, orgulloso, confiable. trabaja duro para sustener a su familia.) in a related commercial, the law office of miguel martinez begins an ad with a working class hispanic couple talking in the kitchen of their home: 1. a. husband: ever since i injured my back, i hurt a lot. without work and without money…? 1. b. (marido: desde que me lastimé la espalda, me duele mucho. ¿sin trabajo y sin dinero…?) 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 3 2. a. wife: we have rights too. 2. b. (esposa: también tenemos derechos.) 3. a. husband: my boss, what will he think of me? and me without papers… 3. b. (marido: ¿mi jefe, que pensará de mí? y yo sin papeles… 4. a. wife: your boss, what does your boss know? let's call the lawyer miguel martinez. 4. b. (esposa: tu jefe, ¿qué sabe tu jefe? hablamos al abogado miguel martinez.) as evidenced by these commercials, a large majority of the intended audience is specifically working class migrants, usually employed in construction or other manual labor. in addition, line three from above overtly addresses the problems of legal status in the country. (“papers” refer to anything from work permits and temporary visas to green cards, any document a person can use to claim a legal right to live and work in the united states.) after the introductory skit transcribed above, the viewer is informed that, “with or without papers,” miguel martinez can help you. these examples, among others not discussed, leave no doubt that the target audience for these commercials is the working-class, monolingual, migrant community. due to the demographics and situations explained above, english is often stigmatized because it is associated with domination and alienation in the larger context of social relations. english, particularly in the public sphere, asserts a hierarchy over spanish and spanish speakers by being the language of authority and directives. spoken spanish may be suppressed externally because it is undervalued, or internally because it calls into question its user’s civic legality. hill (2001) found that spanish use is often seen as unacceptable and threatening in the public realm, and is discouraged (453). many english speakers assume that hispanics who speak spanish in public do so because they do not know english. the rationale behind this being that to become a citizen one must first learn english. thus, questions of legality become dubious when english is absent from the conversation. unfortunately, (perceived) legal status is often another hierarchical determinant that may give rise to other forms of discrimination. this discrimination extends easily to the workplace where, although immigration laws for workers continue to be laxly enforced,1 there is always the latent fear that if one disobeys the boss, immigration officers might be alerted in retaliation.2 for these reasons, i posit that english is less predominant in spanish commercials than might be expected given the hegemony of english in the u.s., because its continuous dominance in public realm makes it undesirable in the private (i.e., in the home.) although english is sometimes still used to market american culture, progress or technology, its overall usage comes much closer to bhatia’s glocalization. that is, u.s. advertisers often choose to adapt to local culture, rather than imposing english on their consumers. my purpose in examining the use of english in u.s.-based spanish language commercials is two-fold. the first task is to deconstruct the idea that english use is homogenous. many studies examine commercialized english by enumerating examples, without completely disambiguating their occurrences. in an effort to correct for this, i 1 since the terrorist attacks of september 11, 2001, the us government has significantly increased security at its borders. however, immigrants already working in the states are largely ignored by immigration officers, unless they commit crimes. 2 this is unlikely to occur since it is unlawful to hire illegal immigrants. many immigrants, however, are unaware of this law or do not believe employers will be penalized. 3 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 begin by examining phonological differences in the pronunciation of english words (i.e. differentiating between english words spoken with an english accent and ones spoken with a spanish accent.) i then provide a discussion of commercials that use entirely spoken spanish and only written english. lastly, to provide a more balanced analysis, i contrast the previous commercials with ones that are devoid of english. after analyzing the data and offering suggestions for the findings at a micro level and within the framework of bourdieu’s (1977) linguistic marketplace, my second task will be to examine the role english plays on a daily basis in a bilingual, macro u.s.-context, and to contrast this with its role in commercials directed at hispanic immigrants. this will be realized in the context of bakhtin’s (1981) heteroglossia as it applies to codeswitching and power relations. urciuoli (1996) and lakoff (1990) will then be introduced into the discussion to help elucidate why advertisers favor authenticating their products, instead of imposing their authority through english use. 2. data 2.1. english with english phonology3 of the seventy ads that employ spoken english, english phonology is used only 40% of the time.4 however, the english in many of these ads resulted solely out of the necessity of pronouncing the product’s name. haarman notes that, “the product name is the element of an ad where a switch into a foreign language occurs most frequently” (in piller, 2003: 172). thus, had there been less products with english names for sale, the occurrence of english would be lower still. although the use of english and english phonology is less common than would be expected based on past studies, a few commercials in which english phonology is found do corroborate many of the previous findings. at times, english phonology is used to index aspects of u.s. culture, technology, fashion and quality. a telling example of cultural stereotypes is a cricket commercial advertising a cellular phone. in this ad, an english rap song plays in the background during the entire commercial. piller notes that, “another ethno-cultural stereotype that english is often associated with is the youth culture, hip hop rebellion, and the street credo of the black urban u.s. ghetto” (2003: 175). in commercials for technology, usually communications, english phonology is always used. (this use is occasionally combined with spanish phonology, but english is never absent.) cosmetics, such as avon and covergirl, which are included in the category of fashion, use english for product names. clothing stores, on the other hand, are less predictable and vary in language use. many companies have separate commercials that fall under 1) english phonology, 2) spanish phonology and 3) written english only. food advertisements are also less consistent: safeway and city buffet seem to use 3 “standard american english (sae)” constitutes the english phonology observed. while this term regrettably has many shortcomings, it is used here to acknowledge a fairly homogenous use of english phonology. for a more in depth discussion, refer to lippi-green (1997) and her descriptions of sae as an “abstraction” and a combination of forms of english, which are not “overly stigmatized” (62). 4 this does not include ads in which both english and spanish phonology were used. a separate section will be devoted to that topic. 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 5 english phonology to index food of a higher quality, while fast-food chains often revert to phonological forms of glocalization. aside from the above usages, however, english phonology more often confirms hierarchical power relations and indexes class mobility. a prime example involves chevy. chevy advertises two trucks, the chevy avalanche and the chevy silverado. the chevy avalanche is a pick-up truck that can be converted into an suv (and vice versa); the ad uses english phonology. the commercial begins with a well-dressed, fairskinned young man driving his (suv) avalanche to meet some friends at a café. as he drives, he notices that he is heading toward the seedy part of town. he checks the address against the closest street sign and, sure enough, this is where he is supposed to meet his friends. a beefy biker in leather watches as he drives by the café. a little shaken, the young man pulls into an alley and quickly converts his suv into a pick-up truck. he then untucks his shirt and musses his hair. as he emerges from the alley, he nods to the biker. when he arrives at the café, his friends are shocked to see him so disheveled. this commercial is aimed at someone who considers themselves to be, or aspires to be, upper-middle class. the dangerous part of town is a place our well-dressed friend has never been. he is clearly out of his element and is not associated with the lower-class. the chevy silverado, on the other hand, is aimed at the working-class; the commercial uses spanish phonology. as the silverado speeds along a rugged mountain road, the person driving the truck is never shown. later, the viewer sees the inside of a motionless truck and the words “power packs” repeat multiple times in spanish phonology. the idea seems to be that the truck and, by extension, the working-class man who drives it (not shown), is tough and powerful. in these two examples it is interesting that the first advertisement (the avalanche) focuses almost entirely on the man's appearance, while the second (the silverado) concerns itself exclusively with the truck's appearance. additionally, it is no coincidence that most silverado models cost about tenthousand dollars less, although upgraded and “fully-loaded” silverados are comparable in price to the avalanche. in these commercials, english and spanish phonology correspond to advertisements for the upper-middle class and the working-class, respectively. the target consumers of the two phonologies are clearly defined. other examples of english phonology that index class mobility include inglés sin barreras (english without barriers, hereafter cited as isb), job corps and my first games. the message of all three is that english skills are essential to success in the united states. isb, a company that specializes in english instruction videos, provides the best example. one ad opens with two men talking: man 1 (m1): what’s going on? hombre 1 (h1): (¿qué pasa?) man 2 (m2): i’m ashamed, but i need to ask you for another loan. hombre 2 (h2): (me da pena, pero necesito pedir otro préstamo.) m1: you can’t keep on like this, man. h1: (usted no puede seguir así, compadre.) m2: yeah, i know i need to learn english but… h2: (sí, yo sé que necesito aprender inglés, pero…) 5 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 m1: here’s your check man, and i’m going to write you another so that you can reserve your english class. h1: (aquí tiene su cheque compadre, y le voy a hacer otro por separado para que reserve su curso de inglés.) m2: thanks, man. h2: (gracias, compadre.) likewise, in all isb commercials, similar examples of mobility and personal success through english proficiency can be found. in addition to the idea that english is the key to success (and spanish is just the opposite), we also find that english with english phonology often reaffirms its authoritative power over spanish. in a related isb commercial, god almighty tells the hispanic man in the commercial to learn english. english authority is also extended to a more mundane level through scenes of the workplace. as discussed earlier, most of the intended viewership for these commercials is working-class hispanics. in a workplace setting, the language of the boss is usually english. thus, english use often connotes domination and authority at work and in society at large. an example of this is another isb ad, which begins with a hispanic worker being demeaned by his hispanic boss. the man impatiently tells his employee in english, "go clean the bathroom! i said, go clean the bathroom!" the fact that this boss is hispanic, but speaks english, addresses an ongoing problem of discrimination against monolingual spanish-speakers by bilingual hispanics. it also suggests that a mastery of the english language is the key to career advancement. later in the commercial, after the employee has learned english, his caucasian boss tells him, “pablo, now that you’ve learned english, you can deal directly with me.” pablo replies, “mr. williams, i thought you’d never ask!” in this case, fluency in english does not earn the man a better job, but it does give him some degree of autonomy and control at work. in addition, and perhaps more importantly, it gives the man a voice which will be heard, understood and respected. the reply, “i thought you’d never ask,” exemplifies the (spanish-speaking) employee’s anonymity and subordination in relation to his english-speaking superiors. it also supports pierre bourdieu’s assertion that in power relations, “some persons are not in a position to speak” (emphasis in the original, 1977: 386). isb and other commercials that use english in this manner hope to promote the idea that through english, the working-class employee can overcome linguistic domination. the need to conform to the dominant culture where english use is the norm is also observed in advertising, namely in terms of bourdieu’s (1977) linguistic devaluation and self-censorship (386). two particularly salient examples from my data involve personal names. many lawyers and individual entrepreneurs pronounce their names with english phonology. this is often true even when the remainder of the commercial is entirely in spanish, and the person’s name is a hispanic one. lawyers yvonne azar, janie castañeda, and john gallegos all pronounce their names using english phonology. english use in this case not only indexes knowledge of english and an american identity (with hispanic ancestry), but might also be a form of self-censorship. many hispanic children who grew up in the united states were (as many children continue to be) denied their language and heritage. spanish use was punished and names were changed to 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 7 english. it is possible, then, that because spanish is devalued in the school system at an early age, the linguistic hierarchy becomes internalized. this internalized devaluation is then manifested as (unconscious) self-censorship, even when speaking spanish to a hispanic audience. the second interesting case is the pronunciation of the english “v”. although this grapheme does exist in spanish, the phoneme does not. thus, in spanish, /v/ is correctly articulated as [b]. in a used-car and a mexican jewelry store commercial, the hispanic announcers both pronounce the letter “v” as [v]. one would expect to find that, especially in the case of the jewelry store, no english would appear at all. typically, stores selling hispanic products attempt to authenticate their goods by omitting english. in fact, the jewelry store conforms to this expectation, with the exception of the [v]. the change is considerably noticeable to listeners because it introduces a foreign phoneme into the spanish discourse. bourdieu sees such speech acts as a sort of hypercorrectness that corroborates an internalized linguistic hierarchy. he writes, “[t]he dominated groups recognize in practice… the legitimacy of the dominant language” (1977: 389). thus, english phonology in spanish discourse is not arbitrary, but rather directly related to societal power relations between the two languages and their respective speakers. 2.2. english with spanish phonology english words pronounced using spanish phonology occurred in 54.28% of the commercials collected. the words are determined to have spanish phonology if they meet any of the four basic criteria listed below:5 1. dropping or softening syllable final—and especially—word final, consonants. 2. shortened pronunciation of vowels (monophthongs are not diphthongized as is often the case in english. also, there is no use of schwa.) 3. substituting english consonants or vowels for spanish ones. 4. use of the spanish sequential consonant structure (i.e., spanish requires an "e" to precede complex onsets. thus, in spanish orthography, the language name is written español and not *spañol.) criterion number one is exemplified when the speaker pronounces walmart as [walmar], instead of [walmrt]. other examples include kmart and lactaid; in both cases the final consonants are also dropped or considerably softened so as to conform to spanish phonology. criterion number two is heard in an ad for the law offices of frickey. the attorney, ms. frickey, whose spanish is a little rusty, nonetheless pronounces her name using spanish phonology (compare with the hispanic lawyers discussed earlier.) the name frickey thus changes from [friki] to [friki]. likewise, crazy leo, a furniture salesman, shortens and raises the “e” in his name to the extent that the name becomes homophonous with the spanish word “lío” (mess, fight). both of these advertisers seem to hope that giving their american names a spanish sound will make them more appealing to hispanic clients. criterion number three is evident in 5 two other criteria have been omitted because 1) they were not observed and 2) because of the difficulties of quantifying them objectively. these are: 1) the substitution of [č] for [š]. 2) use of spanish intonation and resyllabification to match spanish pronunciation. 7 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 a mcdonald’s commercial, in which english vowels are also replaced with their spanish counterparts. the word “chicken” thus becomes [chikεn], in place of [chikn]. with regard to consonants, chevy is pronounced [šeby] instead of [ševy], when it is appropriated with spanish phonology. criterion four is unusual in advertisements6 (it only occurs once in my sample) but is worth mentioning. in burger king's commercial for its new “smoky bbq chicken sandwich,” the client orders his meal saying, “un esmoki bbq, por favor” (“one smoky bbq, please”). in this instance, instead of translating the word “smoky,” it has been appropriated into spanish by adding an “e” before the word. this word, which could have been translated or appropriated with spanish phonology, instead becomes a new spanish word (esmoki). the above examples show how decidedly non-hispanic people and food names are spoken with spanish phonology in an attempt at transforming the products’ images from anglo to hispanic ones. ms. frickey (attorney) and crazy leo (furniture salesman) are perhaps both of hispanic ancestry, although ms. frickey especially, does not possess a superior command of the spanish language. however, it is interesting to note that these people try so hard to erase english from their commercials, while hispanics mentioned earlier used english phonology frequently. the food advertised is, similarly, mentally imbued with hispanic flavor by articulating it in spanish phonology or even creating a spanish word to accommodate its imagery. “smoky” may be an unappetizing or even unintelligible word, but “esmoki” conjures up images of spicy chicken breast roasted on the grill a la mexicana. separate meanings may be mapped onto the same word, depending upon pronunciation and context. while the word is the same, phonological differences create disparate connotations of the interlocutor, which will have a direct bearing on the manner in which speech acts are interpreted. in advertising then, it becomes necessary to address such differences. although [walmart] and [walmar] both mean walmart, each serves to connote a distinct product personality. the act of appropriation calls into question whether or not the word can now be considered solely an english one. i suggest that phonological adaptation blurs a seemingly clear-cut distinction between the two languages, which often results in a lexicon that is more accurately described as spanish than english. the outcome is that, through this advertising tactic, the word may acquire two or more connotations, depending on the manner in which it is pronounced. as such, the english word articulated in spanish phonology ceases to be an english word at all. in this manner, both the word and the product are appropriated through a form of glocalization. 2.3. spanish and english phonology because of the paucity of ads (5.71%) which combine english and spanish phonology, it is difficult to draw many substantial conclusions. nonetheless, the lack of the ads themselves can offer insight into their significance. the first deduction, based on the data as a whole, is that bilingual viewers are not being targeted specifically and that the main audience is comprised of monolingual spanish speakers. a second inference supports the claim that products that use different phonologies actually represent different 6 although unusual in advertisements, this type of speech act is rather common in colloquial discourse. 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 9 products. thus, these few ads may be targeting both hispanic immigrants and chicanos, who are likely to have different conceptions of the same product. 2.4. written english only although a high percentage (54.28%) of commercials that use spoken english have adapted english to spanish phonology, many advertisements forego spoken english altogether. of the 85 ads that use some form of english, 17.65% used spoken spanish and english only occurred in written form. in one instance, an ad for the phone card telecentavos, i propose that a caricature of an english newspaper is used to index hispanics living in the united states (see appendix 1.1). the title of the newspaper, “usa news” effectively targets the hispanic consumer who, because of his/her geographic location, encounters english on a daily basis. the ad does this without losing its hispanic authenticity, because english is never spoken and an archetypal mariachi is shown on the card. the newspaper reads “usa news”, instead of the new york times or the washington post, because these names are less accessible to many immigrants. this same community, however, should have no trouble understanding “usa” and the word “news” will likely be comprehensible as well. nonetheless, the above is a solitary example; written english does not often seem to index life in the united states. more commonly, the occurrence of written english, accompanied by an absence of spoken english, constitutes erasure. with regard to written language in advertising, piller theorizes that in germany: since written discourses tend to be more authoritative… the meaning of english as authoritative is strengthened by presenting it in both the written and spoken modes, while german is only spoken (2001: 160). however, i argue that for an audience who is not literate in english, the spoken language will take precedent over the written. only the necessary vestige of the foreign product's logo remains to speak silently in print. english becomes erased or, at best, concealed. many viewers are unable to understand the words written in english and, thus, these words blend into the background and do not constitute a principle source of information (see appendix 1.2). many of these words are shown out of necessity because they are inherent to the product (i.e. its name.) this is true for the “natural reducer cream” that is included with the la belle corset, among others. in addition, many are presented so quickly that even native english speakers might miss them. aside from these examples, even stronger cases exist to corroborate the idea that written english (without spoken english) is erasure. two jc penny ads illustrate nicely how this works. the two ads are exactly the same, except for the last part. at the end of the first one, the ad says, “we have everything inside, jc penny” (“tenemos todo dentro, jc penny.”) the second ad, however, leaves off the spoken “jc penny,” thus completely eradicating any occurrence of english in the commercial—except for the written name. an even more telling example involves eldercare/meals on wheels. the product is a delivery service that provides pre-cooked meals to the disabled or the elderly; aids patients also commonly use it. the commercial informs us that it is becoming more difficult for mom to care for dad, and that there is a service that can help. however, the words “eldercare” and “meals on wheels” are never spoken. also, the words “meals on wheels” disappear so quickly from view that it is difficult to capture the frame, even in 9 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 slow motion and with a computer program designed to record such images (see appendix 1.3). use of written english in this manner disassociates a stigmatized english product from the product shown. the fact that the commercial for the united states air force only employs written english, as well, fortifies this argument. the air force is necessarily complicated by questions of identity and loyalty. after serving in the armed forces for three years, a person can begin the process of applying for u.s. citizenship. however, citizenship is not always seen as a goal one should aspire to, nor is the military often thought of favorably. apparently having realized this, the air force frames its ad in terms of upward mobility and future success, in addition to not articulating the words “air force.” thus, i argue that, while english does appear in the above commercials, its appearance is not substantial enough to be considered english use. in previous studies, however, the use of english was not thoroughly explored, and the above examples of english appropriated with spanish phonology and instances of written english would have constituted a large part of the ‘english’ data. 2.5. spanish only7 spanish-only commercials accounted for merely 16.66% of the corpus. (however, as i have argued above, many commercials that employ english should actually be considered principally spanish due to appropriation and erasure.) the spanish-only advertisements consist mainly of hispanic products, fortune tellers, lawyers and immigration services. the lack of english authenticates the products as hispanic ones, while simultaneously acknowledging a monolingual-spanish audience. 3. discussion: authority vs. authenticity advertisements are perhaps the prototypical embodiment of linguistic marketplaces where not only products, but also language and culture, are for sale. however, as i have attempted to demonstrate throughout this paper, the fact that english enjoys a dominant hierarchical position in u.s. society does not bequeath it the same legitimacy in hispanic advertising—even within the united states. often, a rejection of ‘proper’ english is an affirmation of hispanic identity and an attempt to reconcile conflicting identities concerning ethnicity or citizenship. marketing tactics concerning english use in u.s. spanish language commercials are perhaps best viewed in terms of bakhtin's heteroglossia, which he defines as: another's speech in another's language, serving to express authorial intentions but in a refracted way. such speech constitutes a special type of double-voiced discourse. it serves two speakers at the same time and expresses simultaneously two different intentions: the direct intention of the character who is speaking, and the refracted intention of the author (emphasis in the original, 1981: 324). the “authorial intentions” in the case of commercials are, of course, marketing strategies 7 “spanish only” does not include the use of english to list an address because no addresses were observed written in spanish. 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 11 aimed at convincing the consumer to buy certain products. the characters, in turn, are the actors who play in the commercials. however, it is the “authors” of the commercials—and not the “characters”—who are of most interest to us here, as it is they who make decisions concerning language use. as discussed earlier, names are a particularly rich site of language contact where code-switching and phonological change are readily observed. product names may be pronounced in english phonology to connote—most often—the product’s ability to offer the consumer class mobility, or even technology, quality or fashion. likewise, product names may be articulated in spanish phonology in an attempt to authenticate them and erase their inherent ‘americaness.’ in either case, advertisements play upon existing stereotypes and value judgments concerning class dominance and ethnicity. whether confirming or rejecting the english language (and the class who speaks it) as dominant and more desirable, a linguistic hierarchy that places english above spanish dictates the framework within which advertisements operate. thus, for the purpose of simplicity, bakhtin’s “refracted authorial intentions” in my corpus of commercials may be reduced to two types. the first type of intention is to affirm, through english phonology, that english is authoritative, product names articulated in english and persons who consume them are superior and in control. the second half of the commercials proposes just the opposite. by articulating english with spanish phonology, or erasing english altogether, advertisements insist that spanish can be used to resist english domination, spanish is the authentic language; hispanic people and products are more desirable. thus, it is through heteroglossia and double-voicing that marketing strategists employ what i term profit-motivated code-switching in order to confirm the authority or authenticity of their products, while using underlying linguistic, ethnic and class ideologies to convince consumers of the veracity of their claims. chart 1. analysis of language use in advertising language use in us-based spanish commercials spanish only 17 ads = 16.66% written english only 15 ads = 17.65% english with spanish phonology 38 ads = 54.28% english & spanish phonology 4 ads = 5.71% english with english phonology 28 ads = 40.00% spoken english 70 ads =82.35% spanish & english 85 ads = 83.33% english only 0 ads = 0.00% 102 for-profit ads commercials 11 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 english english only + english with english phonology + english & spanish phonology: 32 ads = 31.37% spanish spanish only + english with spanish phonology + written english only + english & spanish phonology: 74 ads = 72.54% by analyzing the different methods through which english is used in spanish commercials, one begins to see that english, although prevalent, is not dominant. thus, although it occurs in 83.33% of the ads studied, if one does not consider instances where english is appropriated or erased, the percentage of english use drops to 31.37% (see chart i.) this finding may seem surprising, until one considers the societal framework in which english and spanish operate in the united states. as has been shown, english is the dominant and authoritative language of the public sphere, while spanish is generally relegated to the home. english is often used to subjugate spanish and spanish speakers, whether at work or in society at large. urciuoli (1996) in her study of puerto ricans in the united states describes public and private life in terms of “spheres of interaction.” she writes: spheres are sets of relations polarized by axes of social inequality. one’s inner sphere is made of relations with people most equal to one; one’s outer sphere is made of relations with people who have structural advantages over one (77). similarly, lakoff (1990) recognizes a distinction between public and private discourses. she argues that private discourse is based on, “shared allusions [and] jointly created metaphors…this privacy both creates and utilizes trust, which itself is in turn symbolically connected with intimacy.” public discourse, on the other hand, is “concrete, since participants cannot count on shared allusions… there is no assumption of trust” (129). because english is the language of dominance and the language of the public sphere, its use immediately precludes a relationship of trust and intimacy between the interlocutor and the monolingual spanish-speaking listener. thus, i propose that one of the main reasons for a relative lack of english in spanish commercials is that the broadcasts enter the private realm of hispanic life. although english dominates and controls the public sphere, spanish rules supreme in the home. after being bombarded by english at work and in public, many hispanics look forward to the relaxing and intimate home environment, largely or completely devoid of english. for this reason, many companies have begun to adapt english to spanish phonology— something akin to bhatia’s glocalization—or even to erase english from their advertisements completely, anticipating that this approach will create an intimacy that authenticates their products in the eyes of hispanic consumers. 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 13 4. conclusion based on previous studies, english use in advertising within the united states conforms to some expectations, while straying from others. unlike piller’s study of germany, english is used less frequently in the commercials in my corpus due to class differences between the respective target consumers. in fact, english comes much closer to bhatia’s glocalization, as advertisers attempt to authenticate their products in the ears of hispanic listeners. english words are, more often than not, articulated in spanish phonology, or even erased. this finding shows why occurrences of english in advertising must be disambiguated, as different representations connote disparate meanings, and may even be considered different languages. although english is the dominant language of the united states, because of its very hegemony, it may be viewed as oppressive and is often relegated to a second-place status in advertising. commercials thus operate in the framework of a linguistic marketplace that is based on a broader social marketplace, in which competing ideologies vie for dominance and acceptance. advertisements play upon different stereotypes to sell products; the most common ploy in this case seems to favor a strategy of coercion that authenticates products, rather than forcefully imposing the products’ authority. 13 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 14 appendix 1.1 examples of written english 14 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 15 appendix 1.2 15 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 16 appendix 1.3 16 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/3 doi: https://doi.org/10.25810/dqk2-bb54 between authority and authenticity 17 references bakhtin, mikhail. 1981. “discourse in the novel.” in c. emerson and m. holquist (trans.) the dialogic imagination: four essays, 258-422. austin: university of texas press. bhatia, tej. k. 2002. “globalization or localization? rural advertising in india.” in richard grant and john rennie short (eds.) globalization and the margins, 57-71. new york: palgrave macmillan. bourdieu, pierre. 1977. “the economics of linguistic exchanges.” social science information. 16(6): 645-668. hill, jane. 2001. “language, race and white public space.” in alessandro duranti (ed.) linguistic anthropology: a reader, 450-464. oxford: blackwell publishing ltd. lakoff, robin. 1990. talking power: the politics of language in our lives. new york: basic books. lippi-green, rosina. 1997. english with an accent: language, ideology and discrimination in the united states. london and new york: routledge. piller, ingrid. 2003. “advertising as a site of language contact.” annual review of applied linguistics, 23: 170-183. new york: cambridge university press. piller, ingrid. 2001. “identity constructions in multilingual advertising.” language in society, 30: 153-186. new york: cambridge university press. urciuoli, bonnie. 1996. exposing prejudice: puerto rican experiences of language, race and class. boulder: westview press. 17 steeby: between authority and authenticity published by cu scholar, 2004 colorado research in linguistics 6-2004 between authority and authenticity: english use in spanish-language commercials in the united states nicole steeby recommended citation microsoft word paper_steeby.doc case marking systems in two ethiopian semitic languages colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by weldu m. weldeyesus. case marking systems in two ethiopian semitic languages* weldu m. weldeyesus university of colorado at boulder this paper presents a description of the case marking systems in amharic and tigrinya. the case system is fully retained in the pronominal and determiner systems of both languages. nominal case markings are, however, observed only in definite objects. the languages are nominative-accusative in their case system as can be seen from the pronouns, the definite article, and pronominal affixes which are attached to the verb to show agreement with subjects and definite objects. there is also interaction between the semantic notion of definiteness and object marking, which takes two forms: an object marker attached to the nominal object and an object marking verbal affix. in both cases, it is only if the object is definite that these object markings appear. this is an instance of split p. the two related object-marking phenomena always coexist in both languages. 1. introduction in this paper, an attempt will be made to give a description of the case marking systems in tigrinya and amharic, closely related ethiopian semitic languages. two languages have been considered instead of just one for comparative reasons. amharic and tigrinya exhibit a nominative-accusative case system. nevertheless, it is important to point out that the languages do not have typical nominal case marking except in objects which are definite. the case system is fully retained in the pronominal and determiner (specifically definite article) system. in addition, verbal affixes which show agreement in person, number and gender with subjects and objects help in giving an idea of what the case systems in the languages look like (amanuel, 1998; leslau, 1995; getahun, 1990; and baye, 1987; among others). as a background, it will be shown why the languages have are said to have a nominative-accusative system starting with the pronominal and determiner (definite article) system of the languages. following this, an attempt will be made to demonstrate how verbal affixes are used to show agreement with subjects and objects in transitive sentences involving verbs with typical transitivity features (hopper and thompson, 1980). this is done in order to demonstrate the interaction between object marking and the semantic notion of definiteness which is one of the major objects of this paper. according to hopper and thompson, “transitivity involves a number of components, only one of which is the presence of an object of the verb. these components are all concerned with the effectiveness with which an action takes place” (1980: 251). in the overall discussion, reference will be made to hopper and thompson (1980) since the issues they raised are central to this paper. hopper and thompson identified the * my sincere thanks are due to professor regina pustet, my typology instructor, not only for introducing me to the subject matter of typology in the best possible way, but also for her critical comments in the course of writing this paper. my thanks are also due for the editors of cril for her their helpful comments. 1 weldeyesus: case marking systems in two ethiopian semitic languages published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 following scalar parameters of transitivity: participants, kinesis, aspect, punctuality, volitionality, affirmation, mode, agency, affectedness of object, and individuation of object. it is the last parameter that would be of immediate relevance to this paper since it is concerned with the properties of objects. hopper and thompson (1980: 253) point out that referents of nouns are either less or more highly individuated depending on the properties they have. they argue that proper nouns are more individuated than common nouns, human/animate more than inanimate, concrete more than abstract, singular more than plural, count more than mass, and referential/definite more than non-referential/ indefinite. since the property definite versus indefinite will be the central concern of this paper, it will be discussed in some detail. in connection with this, the authors state, “an action can be more effectively transferred to a patient which is individuated than to one which is not; thus a definite o is often viewed as more completely affected than an indefinite one” (hopper and thompson, 1980: 253). they also posit a hypothesis which they claim to be a universal property of grammars. they call it ‘transitivity hypothesis’: “if two clauses (a) and (b) in a language differ in that (a) is higher in transitivity according to any of the [transitivity] features, …then, if a concomitant grammatical or semantic difference appears elsewhere in the clause, that difference will also show (a) to be higher in transitivity” (1980: 254-255). there are concomitant structural manifestations such as the presence of an object marker or an object marking verbal affix which shows agreement with the object which will be one of the main purposes of this paper. the semantic notion of definiteness is treated in this paper for the reason that it strongly correlates with a split case marking in the coding of objects. in the next sections of this paper, a discussion of the pronominal and determiner systems of the languages will be presented followed by a discussion of the interaction between object marking and definiteness. 2. case systems in tigrinya and amharic both amharic and tigrinya are nominative-accusative in their case system. as pointed out above, there are no nominal case markers except in definite objects. in both languages, case forms are morphologically realised in the pronominal system which has separate paradigms for nominative, accusative and genitive cases. they are also realized overtly in the determiner system of nps, and more specifically the definite article. while the definite article in tigrinya is realized as a separate word ( tfor nominative and n tfor accusative), it is realized in the form of a suffix in amharic (-u for basic nominative and -un for accusative). in tigrinya, there are inflectional paradigms for number and gender in the singular and plural forms. in amharic, however, it is only in the singular that we see gender distinction. 2.1. pronouns tables 1 and 2 below show the pronominal systems of tigrinya and amharic. as can be seen from the tables, in the case of tigrinya, in the nominative forms of the personal pronouns, we can see that the stem n ssis common (except in the first person singular and plural) and the other forms are conjugated using agreement markers. in the 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/4 doi: https://doi.org/10.25810/aq41-pj57 case marking systems in two ethiopian semitic languages 3 accusative forms, n is an object marker, as we shall see in the next section, which appears at the beginning of all the forms and suffixes which show person, number and gender concord are added to it. table 1: tigrinya pronominal system person nominative accusative genitive 1s an n ay nat y/nayy 1p n na n ana natna/nayna 2ms1 n ss xa n axa natxa/natka/nayxa 2fp n ss xi n axi natxi/natki/nayxi 2mp n ss xatum n axatum/ n axum natxatum/natatxum/nayxatum 2fp n ss xat n n axat n/ n ax n natxat n/natatx n/nayxat n 3ms n ssu n u natu/nayu 3fs n ssa n a a nata/naya 3mp n ssatom n atom/n om natatom/natom/nayatom 3fp n ssat n n at n/n n natat n/nat n/nayat n in the genitive, nat(nay-) is the stem common to all the inflected forms in the paradigm. following it are added the inflectional suffixes showing the agreement features of person, number and gender. table 2: amharic pronominal system person nominative accusative genitive 1s ne l ne y ne 1p nna l nna y nna 2ms ant lant yant 2fs anci lanci yanci 2mp 2fp nnant l nnant y nnant 3ms ssu/ rsu l ssu/l rsu y ssu/y rsu 3fs sswa/ rswa l sswa/l rswa y sswa/y rswa 3mp 3fp nn ssu/ nn rsu l nn ssu/l nn rsu y nn ssu/y nn rsu in the case of amharic, as can be seen from table 2, there is more diversity, the third person singular forms being more anomalous than the others. while all other forms 1polite forms exist in the pronominal system of the languages. in tigrinya, there are polite forms for 2ms n ss xum or n ssom, for 2fs n ss x n or n ss n, for 3ms n ssom, and for 3fs n ss n. in amharic, there are also polite forms but with no gender distinction: rs wo for 2ms/2fs and rsac w 3ms/3fs. in tigrinya, the use of polite forms applies to the determiner system as well where we have the same inflections for both the 3mp and its corresponding 3ms polite form on the one hand and for 3fp and 3fs polite form on the other. mason (1996: 20) points out that the polite forms, which are used when talking about or addressing a person one defers, may be used as plurals. 3 weldeyesus: case marking systems in two ethiopian semitic languages published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 involve an alveolar nasal as part of the stem, the third person singular forms do not. the object marker l and the genitive marker y are added as prefixes taking the nominative as a stem to derive the accusative and genitive forms respectively. the presence of alternations in both tables is attributed to dialectal differences in both languages. 2.2. the determiner (definite article) the definite articles in tigrinya and amharic show more differences than the pronominal system. while there is a stem for the article in tigrinya, it is realized in the form of a suffix in amharic as shown in the following tables. table 3: tigrinya definite article nominative accusative examples gloss ms t-i n t-i ti/n ti w di w ddi ‘boy’ fs t-a n t-a ta/n ta gwal gwal ‘girl’ mp t-om n t-om tom/n tom aw ddat aw ddat ‘boys’ fp tn n tn t n/n t n awal d awal d ‘girls’ as can be observed from table 3, tis the stem common to both the nominative and the accusative forms of the definite article in tigrinya. what makes them distinct from each other is the presence of the object marker with the accusative forms. note that the underlying representation for n tis n t-, which has been realized as n tat the surface level2: t→ n t→ n t. table 4: amharic definite article nominative consonants vowels accusative examples gloss ms -u -w -un s w yye-w/-un s w yye ‘man’ fs -wa/-itu/-i( )twa -wa/-yet/-yetwa -un set yyo-wa/-n set yyo ‘woman’ mp fp -u -u -un s wwocc-u/-un setocc-u-/-un s wwocc ‘men’ setocc ‘women’ the amharic definite article shows a more complicated picture as can be seen from table 4. there is no stem representing the definite article as discussed above in the case of tigrinya. it is realized in the form of a suffix attached to the noun it specifies. the 2 this phenomenon of phonological reduction seems to be quite common both in amharic and tigrinya. for example, in tigrinya, the directional prepositions kab ‘from’ and nab ‘to’ are composed of k + ab and n + ab respectively. likewise, in amharic, as can be seen from table 2, there is phonological reduction when the object marker and the genitive marker are added to the nominative, which serves as the stem. for example, in the first person singular, when the object marker l and the genitive marker y are added to the nominative ne, the initial sound of the stem has been deleted yielding the following: l + ne = l ne and y + ne = y ne. 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/4 doi: https://doi.org/10.25810/aq41-pj57 case marking systems in two ethiopian semitic languages 5 amharic definite article is a suffixed element and has different realizations depending on whether the noun to which it is attached ends in a consonant or a vowel, singular or plural, and masculine or feminine. according to leslau (1995: 155), if the noun to which it is attached is masculine singular and ends in a consonant, the marker of definiteness is u as in bet ‘house’, which becomes bet-u ‘the house’. if the masculine singular ends in a vowel, however, the definite suffix is -w as in resa ‘corpse’, which becomes resa-w ‘the corpse’. on the other hand, if the noun to which the definite suffix is attached is feminine singular and ends in a consonant, the marker is realized as -wa, -itu, or -itwa ( twa) used interchangeably as in g r d ‘maid’ which becomes g r dwa/ g r ditu/ g r ditwa ‘the maid’. if the noun is feminine singular and ends in a vowel, the suffixed element is -wa, y tu (-ytu) or -y twa, again used interchangeably, as in doro ‘hen’, which become dorowa/ doroy tu/ doroy twa ‘the hen’ (leslau, 1995). there is no gender distinction in the plural in amharic. the plural marker for all nouns is -occ or -wocc, the former for nouns ending in a consonant and the latter for nouns ending in a vowel, the -wserving as an epenthetic segment between the cluster of vowels which is not permissible in both amharic and tigrinya. the definite marker added to these regardless of whether the noun is treated as masculine or feminine in the singular is -u. for example, in the masculine, as in n gusocc ‘kings’ (the plural of n gus ‘king’), the definite form becomes n gusocc-u ‘kings’. in the feminine, as in n g stocc ‘queens’ (the plural of n g st ‘queen’), the definite form becomes n g stocc-u ‘the queens’. the are certain forms which do not show any gender distinction and hence can be used for both masculine and feminine. one such example is ast nagaj ‘waiter/waitress’. the addition of an appropriate definiteness suffix determines the gender. for example, if -u is suffixed to it, it becomes ast nagaj-u ‘the waiter’; if -wa is suffixed to it, it becomes ast nagajwa ‘the waitress’. in the plural, there is no difference as discussed above: ast nagajocc ‘waiters/waitresses’ becomes ast nagajocc-u ‘the waiters/the waitresses’. it should be noted that while a distinction is made between masculine and feminine in both singular and plural in tigrinya, it is only in the singular that a gender distinction is made in the case of amharic. “for the plural, no distinction is made between the masculine and the feminine” (leslau, 1995: 155). 2.3. pronominal affixes with the above introduction to the pronominal and determiner system of tigrinya and amharic so as to give a picture of the case system, we now see how case forms are indicated with the help of pronominal affixes attached to the verb. this will also help to further show the nominative-accusative nature of the case systems of the languages. tigrinya: 1a) kasa m s’i -u kasa came-3mss ‘kasa came.’ b) kasa anb ssa-tat q til-u kasa lion-pm killed-3mss 5 weldeyesus: case marking systems in two ethiopian semitic languages published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 ‘kasa killed lions.’ amharic: 2a) kasa arr f kasa rested/died-3mss ‘kasa rested/died.’ b) kasa anb ss-oc g dd l kasa lion-pm killed-3mss ‘kasa killed lions.’ the pronominal affixes attached to the verbs in the above examples show that amharic and tigrinya are nominative-accusative in their case systems. in (1) and (2), we see simple intransitive sentences in the (a) examples, and transitive sentences involving a direct object in the (b) examples. while the subjects of the intransitive sentences (s) and those of the transitive sentences (a) share similar features, the objects of transitive sentences (o/p) behave differently. for example, no matter whether the sentence is transitive or intransitive, the pronominal affix which is attached to the verb always agrees with the subject (s or a) in person, number and gender. the presence of a pronominal affix corresponding with the object appears only if the object is definite as will be shown in the following section. 3. object marking and definiteness apart from the typical case marking systems demonstrated above, the two languages also demonstrate a situation in which there is an interaction between object marking and the semantic notion of definiteness. the object marking takes two forms: using an object marker attached to the object or an object marking pronominal affix attached to the verb. 3.1. object marking pronominal affix and definiteness the pronominal affix which marks the object in both languages interacts with the definiteness of the direct object in sentences with monotransitive verbs. put another way, it is only if the nominal object is definite that the object marking verbal affix is attached to the verb. this is one instance of split p system. the split in this case is that if the nominal object is definite, the object marking verbal affix is attached to the verb; if the nominal object is indefinite, the object marking verbal affix is not attached to the verb. tigrinya: 3a) kasa anb ssa-tat q til-u kasa lion-pm killed-3mss ‘kasa killed lions.’ b) kasa n t-om anb ssa-tat q til-u-wwom kasa omdef-3mp lion killed-3mss-3mpo ‘kasa killed the lions.’ 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/4 doi: https://doi.org/10.25810/aq41-pj57 case marking systems in two ethiopian semitic languages 7 amharic: 4a) kasa anb ssoc g dd l kasa lion killed-3mss ‘kasa killed lions.’ b) kasa anb sso-c-u-n g dd l-acc w ( -acc w) kasa lion-pm-def-3mp-om killed-3mss-3mpo ‘kasa killed the lions.’ in examples 3 and 4, the (a) sentences involve indefinite nominal objects, while the (b) sentences involve definite nominal objects. if the nominal object is indefinite, what we observe is that there is no pronominal affix attached to the verb. on the other hand, if the nominal object is definite, a pronominal affix which agrees in person, number and gender with the object is attached to the verb following the pronominal affix which shows concord with the subject. the fact that the pronominal affix which shows agreement with the subject is always attached to the verb is due to the pro-drop nature of the language. pronominal subjects are optionally used unless there is some pragmatic reason to retain them. yet, the verbal affixes attached to the verb always help in identifying the pronominal subject even if it may not appear overtly. in sum, what we see in the above discussion is an interaction between the occurrence of the object marking pronominal affix and the semantic notion of definiteness. it is only if the object is definite that the pronominal affix occurs. 3.2. object marker and definiteness there is also an interaction between an object marker and the semantic notion of definiteness. note that the object marker discussed here is different from the pronominal affix attached to the verb which was discussed in the previous section. the presence of an object marker along with the nominal object (preceding the object in the case of tigrinya and following it in the case of amharic) is likewise determined by whether the object is definite or not. similar to the discussion in the preceding section, this is an instance of split p. in other words, the split would be between the occurrence of an object marker with the nominal object which is definite and its absence when the nominal object is not definite. tigrinya: 5a) elsa anb ssa q til-a elsa lion killed-3fss ‘elsa killed a lion.’ b) elsa had s b ay q til-a elsa one man killed-3fss ‘elsa killed a man.’ c) elsa n t-u s b ay q til-a-tto elsa omdef-3ms man killed3fss-3mso ‘elsa killed the man.’ d) elsa n -kasa q til-a-tto elsa om-kasa killed-3fss-3mso 7 weldeyesus: case marking systems in two ethiopian semitic languages published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 ‘elsa killed kasa.’ amharic: 6a) elsa anb ssa g dd lc elsa lion killed-3fss ‘elsa killed a lion.’ b) elsa and s w yye g dd lc elsa one man killed-3fss ‘elsa killed a man.’ c) elsa s w yye-wn g dd lccw elsa man-def3ms-om killed-3fss-3mso ‘elsa killed the man.’ d) elsa kasa-n g dd lccw elsa kasa-om killed-3fss-3mso ‘elsa killed kasa.’ as shown in examples 5 and 6, when the object is not definite, no matter whether human or non-human, referential or non-referential, the object marker does not appear along with the direct object as in the (a) and (b) examples. nevertheless, the object marker appears along with the object when the object is definite as in the (c) examples and proper nouns, which are inherently definite as in the (d) examples. hopper and thompson mention a situation in spanish which “shows an extreme restriction in requiring that o’s marked with a must be not merely animate, but also either human or human-like–and furthermore that they be referential, as opposed to merely definite” (1980: 256). in amharic and tigrinya, it is only definiteness, but not humanness or referentiality that matters. quoting berman (1978), hopper and thompson discuss that o-markings occur with definite o’s regardless of referentiality or animacy taking an example from modern hebrew, in which an indefinite o is not marked as the object (but unmarked like the subject) while a definite o is marked with the object marker et, in addition to the definite article (1978: 256). this shows that there are two accompanying formal representations an instance of double marking of the definite accusative, i.e., attaching an object marker to the nominal object, and adding a pronominal affix to the verb showing agreement features. in fact, these features seem to coexist all the time, as one does not occur without the other. it is also important to see what picture the object marking takes in sentences involving verbs taking two objects, commonly called indirect and direct, or first and second objects. dryer (1986) calls these primary and secondary objects respectively. the verbal suffixes which show agreement in person, number and gender with the subject and the object are sequentially attached to the verb in order to signal the subject and indirect object agreement features. tigrinya: 7a) kasa n aster m s’ afti hib-u-wwa kasa om-aster book gave-3mss-3fso ‘kasa gave aster a book.’ 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/4 doi: https://doi.org/10.25810/aq41-pj57 case marking systems in two ethiopian semitic languages 9 b) aster n -kasa m s’ af hib-a-tto aster om-kasa book gave-3fss-3mso ‘aster gave kasa a book.’ amharic: 8a) kasa l -aster m s’haf s t’t’-at ( -at) kasa om-aster book gave-3mss-3fso ‘kasa gave aster a book.’ (getahun, 1990: 176/178) b) aster l -kasa m s’haf s t’t’ccw aster om-kasa book gave-3fss-3mso ‘aster gave kasa a book.’ examples 7 and 8 involve ditransitive verbs. whenever there are two objects, the theme argument remains unmarked, while the recipient/beneficiary is doubly marked in both languages: a dative-benefactive marker preceding the recipient argument, and a verbal affix showing agreement with it appears next to the pronominal affix designating the subject. the agent and recipient arguments in the (a) examples, which differ in gender have been switched in the (b) examples. consequently, we see a change in the agreement markers which are attached to the verb. in his discussion of predicates taking two objects, dryer argues, “in transitive clauses containing a notional do but no io, the object affix represents the notional do. in clauses containing both a notional io and a notional do, the object affix represents the notional io” (1986: 812, emphasis added). in his cross-linguistic study of topic, pronoun, and grammatical agreement, givon argues in the same line: “when both an accusative and dative-benefactive are present in the neutral word order, dative agreement takes precedence over accusative” (1976: 162). the above phenomenon partly goes along with the issue of humanness. since the indirect object represents the recipient, which is most likely to be human, it takes precedence over the direct object which represents the theme. hence, the object marking pronominal affixes which agree with the recipient appear next to the subject marking affix, while the direct object remains unmarked. in their discussion of sentences with ditransitive verbs, hopper and thompson (1980) argue that when a human object is in competition with an inanimate object for overt marking, it is the human object that wins. since their study is discourse-based, the generalization that hopper and thompson give makes more sense. although this appears to be the general tendency, there can also be situations in which a non-human object can win. nevertheless, what matters most in the case of amharic and tigrinya is not the humanness/non-human, but definite/indefinite distinction. a similar situation is observed in the case of argument promotion. one such instance is the causative. in both amharic and tigrinya, a sentence with a monotransitive verb becomes ditransitive with the addition of the causative morpheme. tigrinya: 9a) kasa b ggi arid-u kasa sheep slaughtered-3mss ‘kasa slaughtered a sheep.’ b) astern -kasa b gi arid-a-tto 9 weldeyesus: case marking systems in two ethiopian semitic languages published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 asterom-kasa sheep caus-slaughtered-3fss-3mso ‘aster caused kasa to slaughter a sheep.’ amharic: 10a) kasab g arr d kasasheepslaughtered-3mss ‘kasa slaughtered a sheep.’ b) aster kasa-n b g asar dcc-iw aster kasa-om sheep caus-slaughtered-3fss-3mso ‘aster caused kasa to slaughter a sheep.’ transitive verbs are two-place predicates. the addition of the causative marker to transitive verbs introduces a causer argument and the ultimate output seems to be similar to non-causative ditransitive verbs. the causer argument assumes the subject position; the patient assumes the direct object position; while the causee takes the indirect object position. as in the discussion of sentences with ditransitive verbs, the object marker is attached to the causee. in addition, the object marking pronominal affix attached to the verb agrees in person, number and gender with the causee, since the causee takes the position of the indirect object. in connection with this, comrie (1989:177) states that in cross-linguistic study of causative constructions involving transitive verbs, the indirect object seems to be the most justified position for the causee which is extremely widespread across the languages of the world. the verbal affixes also help in signalling the subject and the indirect object in the causative construction just similar to noncausative constructions with ditransitive verbs as stated above with the subject marking pronominal affix appearing immediately next to the verb and then followed by the indirect object marking pronominal affix. the basic object of the monotransitive verb, however, remains unmarked. 4. summary this paper set out with the object of presenting a description of the case marking system in system in tigrinya and amharic, closely related ethiopian semitic languages. while there are no nominal case markings except in definite objects, the case marking system is fully retained in the pronominal and determiner (definite article) systems of both languages. this paper has also shown that the languages are nominative-accusative in their case system, with the help of not only pronouns and the definite article, but also pronominal affixes which are attached to the verb to show the agreement features of person, number and gender with subjects and definite objects taking typical transitive sentences. an attempt has also been made to show the interaction between object marking and the semantic notion of definiteness. object marking has been seen from two perspectives. on the one hand, there is an object marker which is attached to the nominal definite object. on the other hand, there is an object marking verbal affix which is attached to the verb. in both cases, it is only if the object is definite that these object markings appear. this is an instance of split p. the two related object-marking phenomena always coexist. 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/4 doi: https://doi.org/10.25810/aq41-pj57 case marking systems in two ethiopian semitic languages 11 whenever the object is definite, not only is the object marker attached to the nominal definite object, but the object marking pronominal affix is attached to the verb. although there are instances where the occurrence of the object markings is closely related to the human and non-human, animate and inanimate, and referential and non-referential distinctions in several other languages as shown by hopper and thompson (1980), it is only the definite-indefinite distinction which is the essential parameter in the case of amharic and tigrinya. 11 weldeyesus: case marking systems in two ethiopian semitic languages published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 references amanuel, sahle. 1998. sewasiw tigrinya b’sefihu (a comprehensive tigrinya grammar). new jersey and asmara: the red sea press, inc. baye, yimam. 1987 e.c. yeamarinya sewasiw [amharic grammar in amharic]. addis ababa: e.m.p.d.a. comrie, bernard. 1989. language universals and linguistic typology (second edition). oxford: blackwell publishers. givon, talmy. 1976. topic, pronoun, and grammatical agreement. in charles n. li. (ed.) subject and topic. new york: academic press. dryer, matthew s. 1986. primary objects, secondary objects, and antidative. language 62: 808-845. getahun, amare. 1990 e.c. yeamarinya sewasiw [amharic grammar in amharic]. addis ababa: commercial printing press. hopper, paul j. and sandra a. thompson. 1980. transitivity in grammar and discourse. language 56: 251-299. leslau, wolf. 1995. reference grammar of amharic. wiesbaden: otto harrassowitz. mason, john. 1996. tigrinya grammar. new jersey: the red sea press, inc. key to abbreviations used 1,2,3 = 1st, 2nd and 3rd person caus = causative marker def = definite e.c. = ethiopian calendar gen = genitive f = feminine m = masculine o = object om = object marker pm = plural marker p = plural s = singular s = subject 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/4 doi: https://doi.org/10.25810/aq41-pj57 colorado research in linguistics 6-2004 case marking systems in two ethiopian semitic languages weldu m. weldeyesus recommended citation microsoft word paper_weldeyesus.doc speaking in queer tongues: globalization and gay language book review william leap & tom boellstorff (eds.) speaking in queer tongues: globalization and gay language. urbana and chicago: illinois university press. 2004. 288 pages. isbn: 0-252-07142-5. reviewed by aous mansouri “speaking in queer tongues: globalization and gay language” (leap and boellstroff, eds.) focuses on the language approaches that different nonheterosexual groups around the world use in order to identify themselves within their local and indigenous contexts. it also demonstrates how these groups simultaneously attempt to identify as part of a larger, worldwide community of sexual minorities. in the introduction, the editors, william leap and tom boellstorff, state that this is by no means an exhaustive look at all the different cultures around the world. however, the book does an impressive and thorough job with the areas of the world that it does cover, including france, germany, montréal, israel, south africa, new zealand, indonesia, thailand, and various minority groups within the u.s. in this regard the editors state that their aim is “to use a more modestly selected series of essays to show how persons who have same-sex desires, subjectivities, and/or communities mediate and renegotiate linguistic process and product under conditions of the ostensible ‘globalization of gay english’ (p. 4). this quotation reflects yet another of the book’s themes: the authors are concerned with how gay english comes into play within these different groups and how speakers conceptualize its use as indexical of a transglobal society. the book is generally a pleasant and easy read without an abundance of overtly scholastic language, making it accessible to a wider audience. each of the chapters was thoroughly researched, and the authors’ personal involvement within the local cultures enhances the accessibility of the presented material by clearly portraying the local perspective of language use within a global frame. the book’s most obvious shortcoming is noted by the editors themselves. early on in the introduction, they express regret that most of the material is directed primarily towards male same-sex attractions and identities. in the editors’ own words, it is their “hope that the issues raised in this collection will encourage more researchers to examine women’s experiences with gay english globalization and trace the linguistic consequences of those experiences in site-specific terms” (p.5). in the book’s first chapter, denis provencher discusses the influence that american language and media has had on french gay identity. specifically, he examines how the print media’s adoption of gay english lexemes, as well as semantic ideas (such as “coming out”), affects the “frenchness” of gay identity. provencher questions whether “a certain authentic french element […] still persists despite the hegemonic presence of both u.s. language and culture?” (p. 25). provencher then offers an insightful analysis of a magazine, têtu, which largely caters to gay men, analyzing the use of english within its pages. he argues that the discourse reflects a strong french identity despite anglophonic hegemony. for instance, he demonstrates how homosexuals’ use of language in colorado research in linguistics. june 2005. vol.18, issue 1. boulder: university of colorado. © 2005 by aous mansouri 1 mansouri: speaking in queer tongues: globalization and gay language published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) france emphasizes a national identity within the republic. he contrasts this discourse with the development of gay discourse in america, which often carves its niche only by separating the gay community from the larger american public. in the following chapter, heidi minning focuses on the differing functions and roles that english plays for germans who identify as part of a non-normative sexuality group. using examples such as the acronym csd ‘christopher street day’ (which stands for the berlin pride parade named after the street in new york where the stonewall riots occurred), minning illustrates how german speakers code-mix in english for different pragmatic reasons: in print advertisements, in hiv/aids campaigns, and when discussing the politics surrounding sexual minorities. she explains: “when an english expression is selected, it tends to index queer identity in a less ambiguous manner than when either expression is used in a monolingual context” (p. 59). finally, minning argues that english code-switching, code-mixing, and borrowings into german are in fact part of lavender german, a term she uses for a style of speech that is used to index membership in a non-heterosexual community. thus, for minning, speakers do not code-switch between german and english, but rather between german and lavender german. ross higgins discusses the language practices that sexual minority members use in montreal, canada. he does a superb job in bringing to light the different sociopolitical issues that arise in a bilingual community such as quebec, which has had an active history in dealing with language conflicts. the chapter raises numerous questions and issues, including the fact that members of this community must have a deep understanding of both francoand anglophonic culture in order to navigate and understand each culture’s nuances and references to gay culture. higgins also raises the question of whether or not it is even appropriate for members within a bilingual society to have a unified language practice. liora moriel’s chapter talks of a trend within israel whereby the lgbt community (lesbian, gay, bisexual, and transsexual) is turning to english because it “seems to provide not only (the perception of) a gender-neutral language but also access and connection to the (imagined) worldwide lgbt community” (p. 107). because hebrew grammatically marks gender on nouns and adjectives, speakers are forced to reveal their sexual identity when asked about their partner or significant other (an act moriel represents as a courageous undertaking). english, on the other hand, does not mark gender grammatically, making it a more appealing language to those who want to hide their sexual orientation in certain interactive contexts. the second part of moriel’s quotation highlights another function english has for this group of israelis: it indexes their identification into a larger, transnational gay community. william leap turns to language use and sexual identity within postapartheid south africa. because the post-apartheid constitution prohibits any kind of discrimination, including discrimination based on sexual identity, the south african situation is a unique one when compared to other surrounding countries. leap offers multiple examples from print media (ranging from the names of local bars to the linguistic choices in a national gay newspaper), language practice, and 2 2 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/5 doi: https://doi.org/10.25810/7xzn-8s54 book review: william leap and tom boellstorff personal narratives to illustrate the diversity of code-switching practices in the south african context, including code-switching between local languages, such as zulu and afrikaans, and english. not only does this diversity reflect the different attitudes each speaker has towards his/her target audience, the very act of codeswitching is also utilized to illustrate a (perceived) connection to a more global gay culture. david murray then approaches the political struggles that indigenous minorities face within larger euro-centric communities. he examines how such struggles matriculate linguistically in the maori society of new zealand. ambivalence toward the dominant english-speaking majority translates into different linguistic practices, ranging from comfortable acceptance of the english term gay to a rejection of that term through adoption of the maori term takatāpui. diachronically, murray shows us that the original meaning of this maori term was an “intimate companion of the same sex” (p. 169). in recent times, however, the semantics of the term has been extended to include transgenderism and same sex attraction (whether to men or women), thus making it a more inclusive term than the english gay. tom boellstroff discusses “bahasa gay” (p. 182) ‘gay language’, a form of slang used by gay men in indonesia. an interesting aspect of boellstroff’s writing style is his decision to italicize the term gay when referring to men within this group. in doing so, he illustrates that “when [he] speaks of “gay men,” [he] refers to indonesian men who call themselves gay, in other words, not based on [his] own determination of who is gay” (p. 185). this italicization, then, serves as a reminder that the term is used in its indigenous indonesian sense, and not in the american/north atlantic constructed sense. unlike most of the other authors in this book, boellstroff claims that english does not play a significant role within this group. when it does, in addition to creating an alliance with the global imaginings of a non-heterosexual community, the borrowed term assimilates to fit local concepts (of that group). boellstroff then introduces the reader to bahasa gay by exemplifying the basic patterns and ways that this slang is created (mostly) from indonesian words. in contrast to the use of lavender slang in other countries, where language has been utilized by nonheteronormative groups for discretion, boellstroff shows that the emergence of this slang “for secrecy is subordinate to uses linked to belonging that are not always apparent to bahasa gay’s speakers” (p. 189). for example, local media and popular culture have appropriated and assimilated different terms from bahasa gay in ways that conform to local culture. boellstroff is clear in stating that even though the mainstream culture has acquired bahasa gay, it is not necessarily more tolerant towards the identities of bahasa gay speakers. the reader is then introduced to thailand where peter jackson talks about the historical categories of phet, a term that has been roughly translated as meaning gender, sex, or sexuality. jackson provides a history lesson on the different thai identities that have emerged or disappeared within the last century. he supports his claims by using print media as a diachronic resource. we are then shown some of the effects of english on thai vocabularies of sexuality: e.g. gay, tom from ‘tomboy’, and dee from ‘lady’. but jackson also illustrates that these 3 3 mansouri: speaking in queer tongues: globalization and gay language published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) english borrowings are too narrow for the complex identities that pre-existed gay globalization: when members of thailand’s diverse gender/sex cultures have engaged western discourses, they have appropriated foreign categories to local understandings rather than abandoning their society’s dominant gender-based view of human eroticism (p. 227). jackson argues that only gay men and transsexuals identify as part of a larger imagined community. the concept of same-sex female love in thailand has no counterpart in global discourses of sexuality because it is based on gender and not sexuality, as are the previously mentioned toms and dees. in the first of two chapters that focus on minority experiences within the u.s., susana peña discusses the history, culture, and society of gay cuban americans in miami. the author first brings up the clash between latin america’s concepts of gender/sexual identity and its north american counterpart. she explains that in latin american cultures, the “sexual aim” (p. 235), and not the gender of one’s sexual partner, is the deciding factor for one’s identity (i.e. whether one was the active or passive partner). she illustrates how english terms (specifically the term gay) are redefined to fit into the localized gender identities of this latin american gender identity system. for example, a fifty-one-year-old man uses the word gay to refer exclusively to the person playing the passive role in a sexual act. in the end, peña shows us that this common linguistic culture helps unite first and second generation gay cuban americans. in the final chapter, e. patrick johnson describes how african american gay men in the united states carry conflicting black and gay minority identities. this conflict results from the fact that gay english is spoken by mostly eurocentric sexual minorities with a history of inflicting black oppression, while black vernacular english is spoken in an african-american heterosexual culture that has a homophobic history. this chapter’s main point is to illustrate how “through vernacular appropriations of heterosexual domestic tropes, black gay men ultimately resist monolithic notions of blackness and gayness and provide space for community-building and sexual agency” (p.253). these men resort to subverting heteronormative terms that are significant within black culture during everyday speech. a perfect example of this surfaces when a member of the group explains the dual meanings of the term family, extending the semantics of the term to include gay men and the familial bond that this sexual community shares. ultimately, this book does an excellent job at introducing the reader to different (male) same-sex cultures and societies around the world. the concepts of globalization and transnational sexual identity are a recurring theme in several chapters. the authors illustrate how the adoption of (mostly) gay english terms helps speakers from different parts of the world identify with a larger, albeit imagined, world-wide community of sexual minorities. the other achievement the book can boast is that it clearly illustrates that the borrowing of a term does not necessitate the borrowing of the concept behind that term. many chapters reveal 4 4 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/5 doi: https://doi.org/10.25810/7xzn-8s54 book review: william leap and tom boellstorff 5 how cultures and languages appropriate foreign terminology into local concepts of sexual identity, whether it be national identity as conceptualized among bilingual speakers in montreal or the gender-based eroticism of toms and dees in thailand. with everything that it covers, this book should appeal to an audience with a variety of interests, including linguistic anthropology and queer studies. moreover, the editors did an excellent job of making the material accessible to people who have little or no background in any of these fields. hence, i would recommend this book to anyone who has an interest in learning about different cultures of the world and how these cultures interact. aous mansouri university of colorado department of linguistics 5 mansouri: speaking in queer tongues: globalization and gay language published by cu scholar, 2005 colorado research in linguistics 6-2005 speaking in queer tongues: globalization and gay language aous mansouri recommended citation book review what's wrong with ei? or, if it ain't broke why fix it? 1 barshi: what's wrong with ei? or, if it ain't broke why fix it? published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/2 doi: https://doi.org/10.25810/7kd3-at43 3 barshi: what's wrong with ei? or, if it ain't broke why fix it? published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/2 doi: https://doi.org/10.25810/7kd3-at43 5 barshi: what's wrong with ei? or, if it ain't broke why fix it? published by cu scholar, 1994 6 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/2 doi: https://doi.org/10.25810/7kd3-at43 7 barshi: what's wrong with ei? or, if it ain't broke why fix it? published by cu scholar, 1994 colorado research in linguistics 1994 what's wrong with ei? or, if it ain't broke why fix it? immanuel barshi recommended citation tmp.1537747365.pdf.llf8a when phonation matters: the use and function of yeah and creaky voice colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by tamara grivičić and chad nilep. when phonation matters: the use and function of yeah and creaky voice tamara grivičić and chad nilep university of colorado at boulder this paper illuminates the conversational functions of the combination of creaky voice quality and the response token yeah. jefferson (1984) described yeah as an acknowledgement token that also projects “a preparedness to shift from recipiency to speakership” (p. 200). this speaker incipiency is not consistent, though. while yeah is sometimes used to indicate a shift from recipient to speaker, it is sometimes used simply as an acknowledgement token. this difference in function of apparently similar items may be related to token shape. this paper examines several telephone interactions and finds the use of yeah with creaky voice to indicate passive recipiency and either a dispreference to continue the current topic, or a disalignment with the primary speaker. this analysis contributes to the study of phonetics in interactional linguistics. in addition, it supports the notion that token-shape distinctions can account for functional differences within token types. it suggests that phonation or other behavior below the word level may be significant in verbal interaction. 1. introduction interactional linguistics is the emerging endeavor to study traditional linguistic interests, including syntax and phonology, using the tools of conversation analysis (selting and couper-kuhlen 2001). at the level of sound patterns, much work has clustered around analyses of prosody and intonation (e.g. couper-kuhlen and selting 1996), and their functions in conversation. to date, there has been less attention paid to narrower, phonetic analysis of speech sounds1. this paper presents an analysis of the conversational functions of the combination of creaky voice quality and the response token yeah. while response tokens are extensively studied within conversation analysis, voice quality has not been widely considered from an interactional point of view. nonetheless, this paper will suggest that qualities such as creaky voice are available to speakers as resources, and that voice quality may interact with the word yeah to perform a set of conversational functions. within the field of phonetics, creaky voice, or laryngealization, has been described as one of a number of phonation types, or voice qualities. ladefoged (1982) catalogs a variety of phonation types, including modal voicing – the normal vibration of the vocal folds which, according to ladefoged, occurs in all spoken languages – as well as aspiration, murmur, "glottal catch," pharyngealization and laryngealization2. however, in his brief discussion, ladefoged mentions only cases in which these glottal distinctions are phonemic; he makes no mention of the occurrence of such phonation types in languages whose speakers do not systematically utilize or orient to them as distinctive. however, implicit in ladefoged's opening remarks is the suggestion that such varieties might occur 1. however, see, for example, local and kelly 1986, fox tree and clark 1997, bybee and schiebman 1999, and jurafsky et al. 2001. 2. this compares to five features for glottal stricture described in ladefoged (2001): [voiceless]; [breathy voice]; [modal voice]; [creaky voice]; and [closed], the setting for glottal stops. 1 grivic?ic? and nilep: when phonation matters: the use and function of yeah and creaky v published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 in any normal speaker's repertoire. indeed, ladefoged (2001) notes that creaky voice "occurs at the ends of falling intonations for some speakers of english," (125) even though english has no laryngealized phonemes. there is, though, no discussion of the functions of creaky voice in languages where it is not distinctive. other scholars have suggested that creaky voice can have communicative function in english. pittam (1987) suggests that, for australian speakers, creaky voice indexes low solidarity and is associated with male speakers. blount and padgug (1976) describe creaky voice as characteristic of english care-giver speech. duncan and fiske (1977) suggest that, when coupled with low pitch, creaky voice can signal the end of a conversational turn. from an interactional point of view, two works bear particular mention. ogden (2001) suggests that among finnish speakers, creaky voice often co-occurs with syntactic completion, pragmatic completion, and sentence final intonation at the end of a turnconstructional unit (tcu; sacks, schegloff and jefferson 1974). such a combination of potential turn-end markers indicates a complex transition relevance place (ctrp; kärkkäinen, sorjonen and helasvuo, to appear), where a current speaker typically gives way to a new speaker. this use of creaky voice contrasts with glottal stops, which are generally not treated as transition relevant, even when followed by a long pause. furthermore, when creak co-occurs with one or more of these elements, but speaker transition is not affected, trp is retracted by, for example, rushing through the next tcu. to date there are no widely reported findings for english orientation to creaky voice which would compare to ogden's findings in finnish. however, jasperson's (1998) work on repair in english suggests that glottal stop may function in a comparable way in each language. in several types of focus-repairs described by jasperson, a speaker may produce a significant pause after a glottalized cut-off. according to jasperson (personal communication), closure cut-off (which can, under the right conditions, be realized by glottal stop) is routinely used to initiate same-turn repair of the tcu-sofar, and to that extent projects more talk to come (the repair), the continuation of the tcu. silences that may follow closure cut-off, before the resumption of phonation, get interpreted as belonging to the speaker, because she has not brought the turn to possible completion. thus, in english as in finnish, glottal closure is not treated as transition relevant. it remains to be seen whether english speakers treat other glottal strictures, such as creaky voice, as marking transition places. unlike creaky voice, the lexical item yeah has inspired a great deal of writing by linguists. in fact, the shear volume of information precludes a thorough review here. however, two studies that bear on the issues discussed here should be mentioned. jefferson (1984) offers a preliminary analysis of the interactional work which speakers can accomplish through the deployment of acknowledgement tokens mm hm, uh huh, and yeah. according to jefferson, mm hm and uh huh mark "passive recipiency" (202). yeah, on the other hand, is said to mark "imminent speakership" (202); that is, a recipient who produces yeah as an acknowledgement token also projects an assumption of primary speakership. however, jefferson points out that not all tokens of yeah prefigure a change in speakership. what accounts for this variability? the answer is not 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/10 doi: https://doi.org/10.25810/36ed-z920 when phonation matters: the use and function of yeah and creaky voice 3 entirely clear, but jefferson suggests that "token-shape" may account for differences in function of apparently similar tokens. this suggestion leaves open the possibility that phonation or other behavior below the word level may be important to the form and function of acknowledgement tokens. the work of drummond and hopper (1993a) is in some ways a continuation and expansion of work begun by jefferson (1984, 1993), particularly jefferson's suggestion that there is a continuum from the passive recipiency of mm hm, to go-ahead markers such as oh really? to the speaker incipiency marked by yeah. drummond and hopper find oh and okay frequently at the end of tellings, often projecting a change in speaker and/or topic3. the tokens mm hm, uh huh, and yeah all occur earlier in the telling, and prefigure a continuation of the telling. when all three tokens occur during the course of an extended telling, mm hm tends to be realized earliest in the sequence, and yeah latest. further, yeah may signal a shift in speakership, with the participant who utters the token taking over as primary speaker. as jefferson (1984) found, though, yeah can also prefigure a continuation of the telling. 2. data and analyses the data for this study consist of approximately twenty-five minutes of telephone conversations. this may be considered ‘found’ data; it was not recorded by the investigators for the purpose of analysis. instead, phone calls were recorded by men who were ‘teasing’ telemarketers, attempting to keep them on the line for as long as possible with no intent of buying the service advertised. the peculiar nature of these conversations may make it impossible to generalize about much of the behavior recorded. however, since speakers seem not to have any metalinguistic knowledge of their ability to manipulate phonation type (despite the facility of manipulation found by jasperson 1998 and ogden 2001), we assume that the particular phonetic behavior described here is not affected by the nature of the conversation4. the investigators worked with audiotapes of the conversations; the tapes were transcribed and coded for the occurrence of various discourse markers. ‘discourse marker’ was defined to include items such as oh, ok, really, mm, mhm, uh huh, and yeah5. within the transcripts, yeah was the most frequent lexical discourse marker, accounting for 87 of the 260 markers coded. also coded was the occurrence of creaky voice, determined impressionistically (see local 1996). a word was coded for creaky voice when creak was hearable over at least one syllable of the word. distributional analyses (see below) showed that yeah with modal voicing tended to be followed by additional speech much more often than yeah delivered with creaky voice. 3. this is comparable to the go-ahead responses that schegloff (1995) describes in pre-expansion sequences and minimal post expansions. 4. reviewers have also pointed out potential ethical dilemmas related to the use of recorded telephone conversations. this is certainly an issue that researchers should be sensitive to. however, all names and individual identifiers have been suppressed from the data. furthermore, since both the company employing the telemarketers and the customers themselves reserved the right to record the interactions, we have decided to use the data. both federal and state law allow for such recording when, as in this case, at least one party grants consent. 5. for a fuller description of discourse markers, see jucker and ziv (1998). 3 grivic?ic? and nilep: when phonation matters: the use and function of yeah and creaky v published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 conversational analysis was then carried out on portions of the data that featured the word yeah, both with and without creaky voice. (1) example 1. creaky and non-creaky yeah. track_02:45-50 44 m: mean if you only make five calls uh you know five times thirty five that's like ((water stops))(0.80) you kno::w (0.6) m'sorry (0.72) little slow on m'math . it's like= 45 r: =oh, it's like a dollar fifty somethin'\ [(>you know uh<) 46 m: [yeah, it's like a dollar fifty five, 47 r: %ye:[ah,% 48 m: [th]at's for five calls at dollar fifty fi[ve, ] 49 r: [%;ye::ah;%] 50 m: that# that's a lot less then paying like four ninety five a month for you just for making like . you know five ca:lls you know, in example 1, neither instance of creaky yeah is accompanied by other speech. both function as minimal responses and indicate passive recipiency (jefferson 1984). however, we would expect the function of the non-creaky yeah to be different. although it still functions as a minimal response acknowledging the past turn, it signals high speaker incipiency. that is, it secures that the speaker who has just uttered yeah will remain or will become the primary speaker. indeed, this is the case following the noncreaky yeah in example 1. in the interaction so far, m has been the primary speaker, describing a telephone rate plan to r. at line 45, r co-constructs m’s turn by supplying the answer to m’s calculation. so, m’s yeah at line 46 demonstrates his intent to remain the primary speaker. on the other hand, creaky yeah in line 47 has the complimentary function indicating passive recipiency. r does not express interest in taking the floor, but rather he merely gives an indication that he accepts or is following m’s ongoing explanation. similarly, after the creaky yeah at line 49, m continues his sales pitch with no interruption from r. (2) example 2. no uptake. track_02:96-101 96 m: =yeah\ so right now you are probably payin' about three ninety five if ya have qwest for your long distance\ 97 (0.62) 98 m: of course i'm# i'm not ((dishes start up again)) exactly certain cause . i mean unless you know which program you have 99 (0.85) 100 r: ri:ght, (0.27) *%yeah right%*= 101 m: =but like i said right now i can getchyou seven cents no monthly fees or minimums\ example 2 illustrates another case of creaky yeah used to indicate passive recipiency. of particular interest in this example is the sequence organization between the two speakers. although m is providing a space for r to reply -this space is evident in the silences between m’s turns -there is no uptake by r. when r finally produces a reply 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/10 doi: https://doi.org/10.25810/36ed-z920 when phonation matters: the use and function of yeah and creaky voice 5 and m does not resume speaking, r produces creaky yeah signaling his intent to remain a passive recipient. from these examples it appears that creaky yeah functions to mark passive recipiency, not projecting further speech from the speaker who produces it. furthermore, creaky yeah may be seen as an attempt to close out a sequence or discontinue the current topic. as seen in example 1, however, this attempt by the primary recipient to close a topic may or may not be respected by the primary speaker. that is, the attempt to close a sequence is not always successful, since it depends on concurrence of the interlocutors. it is often the case, as in example 2, that creaky yeah displays dis-alignment or dispreference for continuing the sequence. thus, creaky yeah can be seen to function both as a marker of passive recipiency and as a tool to accomplish sequence closings. the following example shows a possibly deviant case. in example 3, which precedes and includes a portion of example 1, m produces creaky yeah and follows it with a substantial turn. (3) example 3. deviant case analysis. track_02:31-44 31 r: [and then . on top ] of the seven cents there [and so was like] forty [two cents] 32 m: [yeah, yeah ] [there . is a . ] there is a thirty-five cents surcharge yeah, 33 r: we::ll there you go, that's what i'm [gettin at, 34 m: [ye:ah,] 35 m: but that's only when you use it i mean you say you don't make many calls i mean you make an an average amount of calls though right? 36 r: i don't know wha[t an average amount of calls is, 37 [((dishes banging continuously))] 38 (0.38) 39 m: you said about five to ten right? 40 r: ((long inhalation)) sss i don't kno:w,= 41 m: =that's about average for most people,= 42 r: =is it? 43 (0.49) 44 m: %ye:ah,% i mean it's# it's not a lot ((water running))of calls really? (0.49) mean if you only make five calls uh you know five times thirty five that's like ((water stops))(0.80) you kno::w (0.6) m'sorry (0.72) little slow on m'math . it's like= in this example, m is the primary speaker. again, he is making a sales pitch, to which r displays recipiency. unlike previous examples, where the primary recipient utters creaky yeah, here m uses the token, at line 44. lines 42-44 constitute an insert sequence within the on-going interaction. at line 42, r produces a checking question, which m answers at line 44 with creaky yeah before resuming the larger activity in which he has been involved. thus, the creaky yeah at 44, followed by sentence final intonation, accomplishes much the same function as that displayed in examples 1 and 2. in terms of speaker incipiency, creaky yeah appears to be similar to mm hm or uh huh, as described by jefferson (1984, 1993) and drummond and hopper (1993a, 1993b). that is, a participant who utters either mm hm, uh huh, or creaky yeah, continues to be a 5 grivic?ic? and nilep: when phonation matters: the use and function of yeah and creaky v published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 recipient, and allows her conversational partner to continue as the primary speaker. there appears to be a slight difference between these tokens in terms of alignment, however. example 4 shows that the use of uh huh indicates alignment with the primary speaker. (4) example 4. alignment with uh huh track_04:127-135 127 r: so thirty five cents no matter what, 128 m: right thirty five cents no matter what\ [and then ( )] 129 r: [and then seven cents] on top of that, 130 m: right [but you know ] 131 r: [so if i ] if i just call one minute\ 132 m: uh-hmm? 133 r: it's seven= 134 m: =forty two,= 135 r: =seven cents plus thirty five, here, r is checking his understanding of the sales-pitch-so-far. at line 128, and again and at 131, m has produced speech in overlap with r. in both instances, m drops out, and allows r to speak. m’s uh hmm? at line 132 indicates passive recipiency, allowing r to continue as primary speaker. further, uh hmm indicates alignment with the turn that r is in the midst of producing. this alignment is further evidenced by the co-construction of the number and its significance in lines 133-135. contrast the alignment shown by uh hmm with the dis-alignment and topic transition that creaky yeah marks in example 5.6 (5) example 5. disalignment with creaky yeah. track_03:358-368 358 m: [yaa::h, . heh, ] =yaa::h\ that's a good college, 359 r: you know i tell you ma:n\ every weekend\ those# those da:mn kids are up there burnin' the da:mn hill down\ 360 r: you [know, kickin' in the windo:ws] 361 m: [he he he ] 362 r: and drinking bee:r and throwing up and all over the damn street\ and [burnin'] sofas/ 363 m: [hhhhh ] 364 (0.598) 365 r: those {expletive deleted} crazy over there those ba:stards, 366 m: %yaa::h%, 367 (0.694) 368 m: so like what you think about that other pla:n, (unintelligible) the surcha:rge with thirty five cent connec[tion fee:? ] in response to m’s observation in line 358, ‘yeah, that’s a good college,’ r produces a telling that may be seen as disagreeing. r describes some negative aspects of college 6. a potentially offensive expression has been removed from example 5. 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/10 doi: https://doi.org/10.25810/36ed-z920 when phonation matters: the use and function of yeah and creaky voice 7 life, and ends with an unflattering characterization of the students, which includes an offensive racial epithet. this remark is followed by m’s creaky yeah at line 366. the use of creaky yeah signals m’s desire for r to end the current topic, but allows for r to remain the primary speaker. following the creaky yeah there is a significant silence, during which r fails to resume speakership. after a pause, m self-selects and begins a new topic. 3. results the data from this pilot study revealed a high proportion of creaky voice tokens occurring on the word yeah (n=5 of 8 creaky voice tokens). as mentioned on the previous section, we conducted a distributional analysis on the audio data comparing occurrences of creaky and non-creaky yeah. tables 1 and 2 show the results of this analysis. the above tables show that instances of non-creaky yeah are likely to be followed by additional speech. this finding is analogous to drummond and hopper’s (1993b) observation that 46% of yeah tokens are followed by speech. the instances of creaky yeah, on the contrary, tend not to be followed by speech. in fact, they appear to show preference to occur alone. speech final non-final total count 37 56 93 % 40% 60% table 1. number of speech-final occurrences of yeah versus yeah followed by speech. speech final non-final total count 4 1 5 % 80% 20% table 2. number of speech-final occurrences of creaky yeah versus creaky yeah followed by speech. 7 grivic?ic? and nilep: when phonation matters: the use and function of yeah and creaky v published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 although our research is preliminary and based on a very small set of data, it nevertheless shows a clear distinction between creaky and non-creaky yeah and functional interaction between voice quality and lexeme. this distinction is not unlike the one presented in jefferson’s (1984) analysis, which showed a distinction between yeah and mm hm. jefferson suggested, “this systematic distinction, raised in a single-instance analysis which generated a collection, can now serve as a resource to be turned to further single-instance analysis, where some otherwise obscure interaction-bits can be brought to focus” (206). our analysis shows that voice quality may interact with words to perform a set of conversational functions. discourse uses of the combination of creaky voice quality and response token implicate the following semantic functions: passive recipiency, a dispreference to continue the current topic, or a disalignment with the primary speaker. these findings support the notion that token-shape distinctions can account for functional differences within token types and that qualities such as creaky voice are available to speakers as resources. our research contributes to the study of phonetics in interactional linguistics as it pays attention to the narrower, phonetic analysis of speech sounds. in addition, it suggests that phonation or other behavior below the word level may be significant in verbal interaction. 4. conclusion this study set out to illuminate the functions of a response token, yeah, and a phonation type, creaky voice. we have demonstrated that creaky yeah is not identical in function to yeah with modal voicing or other types of glottal stricture. while jefferson (1984) suggests that yeah generally signals high speaker incipiency, yeah in conjunction with creaky voice signals passive recipiency. this observation may support and explain jefferson’s suggestion that token-shape distinctions can account for functional differences within token-types. considerable work exists to describe the functions of yeah. a recurrent suggestion (jefferson 1984, 1993; drummond and hopper 1993b, 1993c; gardner 1998, 2001) is that yeah marks speaker incipiency. to date, however, analysts have described this incipiency as variable. while yeah can both respond to a previous utterance and project continued speech, it is not always followed by further talk. the present analysis suggests one possible reason for this: creaky yeah features speaker incipiency so low that it has the complementary function of indicating recipiency. very little interactional research has been done on creaky voice. ogden’s (2001) research on glottal phonation and turn transition in finnish is one example of such work. ogden’s suggestion that creaky voice occurs at the ends of turns, and signals transition relevance, is compatible with our findings. in finnish, creaky voice signals the end of a speaker’s turn. this invites an interlocutor to take the floor, and projects recipiency on the part of the speaker who produces creaky phonation. we have suggested that, in english, creaky yeah similarly signals recipiency and requests a change in topic. our further suggestion that creaky yeah indicates dis-alignment or dispreference may be limited to english or to this token alone. it remains to be seen what, if any other functions creaky voice has in english talk in interaction. 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/10 doi: https://doi.org/10.25810/36ed-z920 when phonation matters: the use and function of yeah and creaky voice 9 appendix a transcription conventions ? terminal rise , terminal fall / non-terminal rise \ non-terminal fall . short pause (< 0.2) (0.6) pause in seconds # cut off or interruption (( )) transcriber’s notes ( ) transcriber’s best guess (unintel) unintelligible speech * * low volume ; ; low pitch % % creaky voice > < relatively fast speech bastards relatively high volume pho:ne long segment ! alveolar click hh exhaled breath nb: these transcription conventions were designed for easy reproduction in ascii character sets. thus, they can be used with most transcription software and most email or other file-sharing software. note also that, unlike earlier systems, each mark of punctuation has only one function. 9 grivic?ic? and nilep: when phonation matters: the use and function of yeah and creaky v published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 references blount, ben and elise padgug. 1976. ‘mother and father speech: distribution of parental speech features in english and spanish.’ papers and reports on child language development 12: 47-59. bybee, joan and joanne scheibman. 1999. ‘the effect of usage on degrees of constituency: the reduction of don't in english.’ linguistics 37 (3): 575-596. couper-kuhlen, elizabeth and margret selting. 1996. prosody and conversation. cambridge: cambridge university press. drummond, kent and robert hopper. 1993a. ‘acknowledgment tokens in series.’ communication reports 6 (1): 47-53. drummond, kent and robert hopper. 1993b. ‘back channels revisited: acknowledgment tokens and speakership incipiency.’ research on language and social interaction 26 (2): 157-177. drummond, kent and robert hopper. 1993c. ‘some uses of yeah.’ research on language and social interaction 26 (2): 203-212. duncan, starkey and donald w. fiske. 1977. face to face interaction: research methods and theory. new york: wiley. fox tree, jean e. and herbert h. clark. 1997. ‘pronouncing “the” as “thee” to signal problems in speaking.’ cognition 62: 141-167. gardner, rod. 2001. when listeners talk: response tokens and listener stance. amsterdam: john benjamins. gardner, rod. 1998. ‘between speaking and listening: the vocalizations of understanding.’ applied-linguistics 19 (2): 204-224. jasperson, robert. 1998. ‘repair after cut-off: explorations in the grammar of focused repair of the turn-constructional unit-so-far.’ ph.d. dissertation, university of colorado at boulder. jefferson, gail. 1984. ‘notes on a systematic deployment of the acknowledgement tokens “yeah” and “mm hm.”’ papers in linguistics 17 (1-4): 197-216. jefferson, gail. 1993. ‘caveat speaker: preliminary notes on recipient topic shift implicature.’ research on language and social interaction 26: 1-30. jucker, andreas and yael ziv discourse markers: descriptions and theory, 171-201. amsterdam: john benjamins. jurafsky, daniel, alan bell, michelle gregory, and william d. raymond. 2001. ‘probabilistic relations between words: evidence from reduction in lexical production.’ in joan l. bybee and paul j. hopper, eds., frequency and the emergence of linguistic structure. amsterdam: benjamins. kärkkäinen, elise, marja-leena sorjonen and marja-liisa helasvuo. to appear. ‘discourse structure.’ in t. shopen, ed., language typology and syntactic description. ladefoged, peter. 1982. ‘the linguistic use of different phonation types.’ university of california working papers in phonetics 54: 28-39. ladefoged, peter. 2001. a course in phonetics. fort worth: harcourt college publishers. 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/10 doi: https://doi.org/10.25810/36ed-z920 when phonation matters: the use and function of yeah and creaky voice 11 local, john. 1996. ‘conversational phonetics: some aspects of news receipts in everyday talk.’ in elizabeth couper-kuhlen and margret selting, eds., prosody in conversation: interactional studies. cambridge: cambridge university press. local, john and john kelly. 1986. ‘projection and 'silences': notes on phonetic and conversational structure.’ human studies 9: 185-204. ogden, richard. 2001. ‘turn transition, creak and glottal stop in finnish talk-ininteration.’ journal of the international phonetic association 31(1): 139-152. pittam, jeffery. 1987. ‘listeners' evaluations of voice quality in australian english speakers.’ language and speech 30 (2): 99-113. sacks, harvey, emanuel schegloff and gail jefferson. 1974. ‘a simplest systematics for the organization of turn-taking in conversation.’ language 50: 696-735. schegloff, e. a. 1995. ‘sequence organization.’ manuscript. selting, margret and elizabeth couper-kuhlen. 2001. studies in interactional linguistics. amsterdam: john benjamins. 11 grivic?ic? and nilep: when phonation matters: the use and function of yeah and creaky v published by cu scholar, 2004 colorado research in linguistics 6-2004 when phonation matters: the use and function of yeah and creaky voice tamara grivičić chad nilep recommended citation microsoft word paper_grivicic_nilep.doc avoiding the wrath of the thunder beings: restricted language and lakhota metaphorical structure 1 catches and gómez de garcia: avoiding the wrath of the thunder beings published by cu scholar, 1994 2 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/5 doi: https://doi.org/10.25810/srnr-we55 3 catches and gómez de garcia: avoiding the wrath of the thunder beings published by cu scholar, 1994 4 colorado research in linguistics, vol. 13 [1994] https://scholar.colorado.edu/cril/vol13/iss1/5 doi: https://doi.org/10.25810/srnr-we55 colorado research in linguistics 1994 avoiding the wrath of the thunder beings: restricted language and lakhota metaphorical structure violet catches jule gómez de garcia recommended citation tmp.1537748506.pdf.hjbpl quoting the unspoken: an analysis of quotations in spoken discourse colorado research in linguistics. june 2007. vol. 20. boulder: university of colorado. © 2007 by jessie sams. quoting the unspoken: an analysis of quotations in spoken discourse jessie sams university of colorado at boulder in conversational speech, quotations are often used by speakers to portray events and stories that have happened in the past to their recipients. however, quotations can also be used to communicate thoughts or ideas that have never been spoken aloud before they were quoted. for instance, speakers can quote themselves by quoting inner thoughts that occurred to them during a particular situation; also, speakers can talk about future events and quote material that has never been (and most likely never will be) said in particular situations. this paper explores these types of quotations to find their uses, possible prosodic cues, and recipient participation.1 1. introduction in conversational speech, quotations2 are often used by speakers to portray events and stories that have happened in the past to their recipients. in the following example, jemma is reporting to her friend ivy an incident that had occurred earlier that day (a key to the transcription is included in the appendix). (1) ‘g’ run-in 1 jemma: um (0.7)so yeah when he told me that i just looked 2 at him and i’m like (0.9) <> are you serious? 3 ivy: ((laugh)) 4 jemma: he was all <> yeahlike give me a call 5 (1.1) 6 ivy: you’re like <> i don’t play that game 7 anymore with you ((laugh)) 8 jemma: i know but i’m just going (0.2) um(0.7) so: 9 <> satur[day huh? 10 ivy: [((laugh)) 11 jemma: <>and if i call you on saturday, you’re 12 actually going to do something about it? (1.2) he 1 i would like to thank everyone who helped me with this paper: barbara fox for getting me interested in conversation analysis, laura michaelis for suggesting working with spoken quotatives and quotations, and the anonymous reviewers who graciously provided helpful comments for editing the paper. also, thank you to my family for all their support. 2 in this paper, all directly reported speech will be referred to as quotations or quoted material. other researchers refer to this as “direct reported speech” or “drs” (e.g. holt 2000, holt 1996, niemelä 2005). 1 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 2 13 looks at me he’s all <> yeah like (0.7) like 14 he looked all hurt [like i hurt his feelings or 15 ivy: [oh my gosh 16 jemma: something (0.5) <> and i’m going <> 17 have you not remembered what [has happened 18 ivy: [((laugh)) 19 jemma: the past like (0.5) year and a half? 20 ivy: aw 21 (0.8) 22 jemma: so((laugh)) so: yeah i just (0.6) i kinda give 23 him this incredulous look like <> what the 24 hell are you talking about? (0.8) and um (1.0) 25 yeah, it just 26 (0.5) 27 ivy: i’m so proud of you 28 (0.4) 29 jemma: i was like <> yeah okay 30 ivy: ((laugh)) 31 jemma: and i walked away going <> yeah i don’t 32 have your number 33 ivy: ((laugh)) 34 jemma: i’m sorry i couldn’t call 35 ivy: mm 36 (0.7) 37 ivy: <> i lost my phone this exchange provides several examples of quotations; two features that distinguish these quotations from indirectly reported speech are the use of the “original’s” deixis and the use of a quotative (holt 1996). for example, in line 2, jemma uses the pronoun you in the utterance are you serious. it is clear, though, that she is not addressing this question to her recipient, ivy. instead, the quoted material shows what she said in the original conversation, and the you refers to the other original participant (who is here simply called ‘g’). therefore, the original deixis remains in the quotation. also, are you serious is prefaced by the quotative i’m like. while this particular instance utilizes both deixis and a quotative, often quotatives can be deleted. jemma is able, through the use of quotations, to not only tell ivy what happened and what was said but also to provide commentary on the situation. the use of prosodic cues (e.g. louder, high, slower) allows jemma to include within the quotation how she feels about the reported situation. there has been a great deal of research on the use of quoted speech and affect (besnier 1993, clark and gerrig 1990, günther 1998, buttny 1997, couper-kuhlen 1998, holt 1996, holt 2000). using quotations allows speakers to have “subtle and intricate ways in which [they] can comment on the utterances they report while simultaneously appearing to simply reproduce them” (holt 2000). in reproducing quoted material, though, speakers are not often expected to be providing a verbatim recall of what was originally said. in their paper “quotations as demonstrations,” clark and gerrig (1990) argue that “quotations are a type of demonstration. just as you can demonstrate a tennis serve, a friend’s limp, or the 2 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 quotations in spoken discourse 3 movement of a pendulum, so you can demonstrate what a person did in saying something” (764). they state that quotations serve as replications of what the speaker wants to convey to the recipients. in other words, quotations are not exact replicas of what was actually said; rather, quotations demonstrate what another person said, how that person said it, and how the current speaker feels toward what was said. clark and gerrig also state that quotations have two functions. the first is detachment: “when [speakers] quote, they take responsibility only for presenting the quoted matter–and then only for the aspects they choose to depict” (792). and the second function is direct experience: “when we hear an event quoted, it is as if we directly experience the depicted aspects of the original event” (793). the quotations jemma provides in lines 2, 4, 8-9, 11-12, and 29 fit in well with clark and gerrig’s idea of demonstrating a previous speech act; however, not all quotations are used strictly for reporting something that was actually said. some quotations are not used as demonstrations of previous speech acts; for example, in lines 16-17 and 19, jemma shows how she feels about the situation with ‘g’ by saying the quotation have you not remembered what has happened the past like year and a half in a much louder voice than the surrounding material. however, when this is taken into context with the rest of the conversation, it is understood that jemma did not actually utter these words; instead, this quotation is provided as a demonstration of how she felt during the situation. therefore, the function of this quotation is different from the earlier examples. it is still quoted material, which can be seen by the deictic use of you and the use of the quotative i’m going; however, it is quoted material that has no original referent. jemma repeats this type of quotation in lines 23-24, 31-32, and 34. ivy, however, uses quotations that are not demonstrations of what she knows to have happened because she has never heard this story before and, therefore, does not know how the story goes. in lines 6-7, ivy aligns herself with the story by adopting the voice of jemma in the story by saying i don’t play this game anymore with you. although this quote was never actually spoken by the original participant in the story, ivy shows she understands the story so far by “chiming in.” holt (2000), couper-kuhlen (1998), szczepek (2000), and niemelä (2005) talk about this phenomenon as a way for the recipient to not only show comprehension of the ongoing dialogue but to actually become a part of it and comment on it. ivy does the same in line 37 when she once again adopts the voice of jemma and says i lost my phone. in both of these instances, ivy is using the deictic pronoun i to refer to jemma–the original participant in the story–and not to herself. by aligning herself with jemma’s stance of the story, ivy has become an active participant in the story rather than an inactive recipient. 3 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 4 in some cases, speakers can provide quotations while telling a story even when they were not an original participant. in the following example, ivy and mary are talking about a story in the past; although neither of the speakers was actually a participant of the story, mary uses quoted speech as it was reported to her: (2) mountain climbing 1 ivy: [((laugh)) 2 mary: [as a matter of fact when um we were out there for 3 christmas 4 (0.3) 5 ivy: oh uh-huh 6 (0.3) 7 mary: my dad took him to um (0.2) mount herman 8 (1.1) 9 ivy: [<> uh? 10 mary: [that littleit’s a little mountain 11 (0.2) 12 ivy: okay 13 mary: out there (0.6) near monument 14 ivy: okay 15 (0.4) 16 mary: and um they went mountain cmountain hiking 17 ivy: oh nice 18 mary: and ((laugh)) and my dadit was just him and my 19 dad (0.6) and um (1.1) they were walking and my dad 20 (0.3) i mean they had walked for a while and my 21 dad’s like (0.6) <> well you ready to turn 22 around? (0.4) and he kept asking him over and over 23 because he knew (0.8) you knowi mean he’s a 24 little kid 25 ivy: right 26 mary: and he would get tired [(x) 27 ivy: [and like <> i’m not 28 carrying you back kid ((laugh)) 29 mary: yeah 30 i & m: ((laugh)) 31 mary: and he’d have to walk all the way back and jason’s 32 like (0.6) <> now you can keep asking me but 33 i’m just going to keep telling you no 34 ivy: ((laugh)) <> nice 35 mary: and so my dad thought he would get slick (0.7) and 36 he made a cut (0.7) you know to start going back 37 but he thought that jason wouldn’t notice 38 ivy: uh-huh 39 (0.5) 40 mary: and jason’s like (0.9) <> papa mike (1.3) 41 we’re not going up towards the mountain anymore 42 i & m: ((laugh)) 43 mary: [my dad 44 ivy: [(x) <> kid we got to the top okay? 45 mary: my dad about died 46 ivy: ((laugh)) 47 mary: and just a couple of minutes later (0.9) jason’s 48 just like (0.9) i gotta take a bri gotta take a 49 rest 50 (0.2) 51 ivy: uh yeah ((laugh)) [he’s like ((groan)) 52 mary: [and he did have to end up having 4 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 quotations in spoken discourse 5 53 to carry him back and i’m telling you whati have 54 never (0.4) seen that baby more wiped out in my 55 life in telling this story, mary uses quotations to demonstrate what she believed to have happened and to have been said on this mountain hiking expedition; even though she was not personally there, she can tell the story with quotations because she has most likely heard the story from the two participants, her dad and her son. ivy, though, has never heard this story and again aligns herself with the story by adopting the voice of the dad in lines 27-28 (i’m not carrying you back kid) and line 44 (we got to the top okay?) and the voice of jason in line 51 with the groaning sound. this example of alignment is different from the previous example because instead of aligning herself with the speaker, ivy aligns herself with the original participants in the story (neither of which are present for this conversation). are these examples of quotations still considered demonstrations even though the speaker is not quoting a known utterance? these types of quotations seem belong to a separate class of quotations, one in which speakers are not demonstrating something they know to have happened; rather, speakers feel free to quote speech that has never been spoken aloud. based on contextual clues and prosodic information, these quotations can be considered demonstrations of mental states of the speaker (e.g. inner speech) or proposed speech for participants in stories that take place in the future. this paper examines the use of quotations that demonstrate inner speech and future dialogue. 2. inner speech nearly everyone has had the experience of listening to a friend recount a story in the past when she quotes herself as saying something completely out of character. and then the recipient has to ask, “did you really say that?” many times, when people tell stories, they include their inner thoughts as reported speech; reporting this unspoken speech gives speakers the chance to “retell” the story so that they can say that witty comeback that they wished they could have actually thought of while the story was taking place. this gives the speaker a chance to put her own spin on the reported dialogue. it also gives her a chance to allow other people to know what she is or was thinking. in this first dialogue, ivy is calling mary for the first time in a while so that they can catch up: (3) took so long 1 mary: hello? 2 (0.4) 3 ivy: mary? 5 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 6 4 (0.2) 5 mary: uh-huh? 6 (0.1) 7 ivy: it’s ivy 8 mary: hi, how are you? 9 ivy: i’m good, how are you? 10 mary: <> i’m good 11 ivy: i’m so sorry it took me so long to get back 12 (0.1) 13 mary: it’s okayi know we all [lead busy lives 14 ivy: [((laugh)) 15 ivy: i ((laugh)) [well i should have written it down 16 mary: [(x) 17 mary: it took me a while to get back to you and i’m like 18 <> goshshe’s going to think i’m horrible 19 ivy: oh i didn’t think that mary quotes her own inner thoughts in lines 17-18 by using a quotative (like) and by using breathy speech for the quotation. here it becomes obvious that this is a quotation and not simply a declaration by mary’s use of the pronoun she; the she mary is referring to is ivy, who is her recipient. if mary wanted to simply say this sentence without quoting it, she would have used the pronoun you since she is addressing her recipient. the breathy quality of her quotation could mean that mary is portraying that she feels badly for not having called ivy sooner; therefore, while this could be considered a demonstration of mary’s mental state, it is not a demonstration of a previously spoken quote. ivy uses these same methods to quote her inner speech while looking around her parent’s house at everything they will have to pack before moving: (4) packing 1 ivy: but it’s just like, i mean there isyou just look 2 around my parent’s house and there’s just stuff 3 [everywhere 4 mary: [<> yeah oh i hate that 5 ivy: you know? 6 mary: yup 7 (0.9) 8 ivy: so i was just like <> oh i can’t even 9 imagine 10 mary: ((laugh)) [and it’ll take [(x) forever 11 ivy: [(x) [((laugh)) 12 ivy: it’s gonna take like a couple months 13 mary: ((laugh)) 14 ivy: it’s like you should just start packing now 15 mary: ((laugh)) in lines 8-9, ivy uses a lower pitch and breathy voice to quote her inner speech (oh i can’t even imagine) while thinking of the daunting task ahead of her parents; again, this demonstrates ivy’s mental state at the time that the story took place. along the same lines, in line 14 ivy again quotes her inner speech (you should just start packing now), but this time her use of normal pitch and the accent on the word now portray a sense of urgency rather than a sense of being overwhelmed. 6 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 quotations in spoken discourse 7 one way a speaker can signal that the quoted material is not something that has been stated aloud is the use of the quotative it’s like (instead of using i’m like). by the use of it, ivy is able to signal that she is simply using a quotation to portray her stance towards the situation. within the quotation, though, she still uses the deictic you to refer to her parents and not to her recipient. in the next interaction, ivy and mary are discussing the gym that ivy has joined; this example differs from the previous two because mary uses the same techniques to give voice to what she assumes to be ivy’s inner thoughts in order to become an active participant (line 18). (5) gym talk 1 mary: well (0.3) uh what gym is it? 2 (0.6) 3 ivy: um it’s better bodies (0.9) it’s a (0.2) it’s not a 4 chain wellit kind of is a chain here in denver 5 (0.4) like there’s three of them (0.6) but it’s not 6 like a (0.2) national chainit’s not like 7 <> 24-hour fitness or anything 8 (0.4) 9 mary: i mean is it a male and female club? 10 ivy: oh yeah (0.1) mm-hmm (0.9) yeah and i like it just 11 well one it’s like (0.3) on my way home from work 12 so (0.7) like i have to pass it every day 13 [((laugh)) 14 mary: [((laugh)) and if you don’t go you feel really bad 15 (x) 16 ivy: and when ias soon as i pass it and i’m going 17 straight home, yes, i feel really guilty ((laugh)) 18 mary: you’re like <> gawd all of these examples show how affect can be portrayed by the use of a quotation even when the quotation is not demonstrating any previously uttered material but is rather demonstrating a thought. 3. future dialogue speakers can also create fictive worlds of dialogue for future situations; because this is dialogue based on speech that has not or probably will not occur, it seems to be a type of “fake” quotation in that the quotations only exist in the world created by the current speakers. even if the situation being talked about is something that does actually occur in the future, the dialogue that was created to go with it will often times not be a part of that ongoing situation. for example, in the next exchange, ivy and mary are talking about their future 10-year high school reunion; this is a reunion that will take place in the future, but the dialogue they propose for this future situation most likely will not: 7 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 8 (6) high school reunion 1 mary: you know what (0.5) and sandra beats herself up all 2 the time she’s like (0.4) <> i’m gonna go 3 back to the high school reunion and i haven’t 4 accomplished anything (0.4) here i am divorced, 5 i’ll have no kids 6 ivy: ((laugh)) 7 mary: and i haven’t done nothing with my life (0.2) i’m 8 like (0.4) whatever 9 (0.4) 10 ivy: aw 11 (0.5) 12 mary: ((laugh)) 13 (0.5) 14 ivy: i’lli’ll be proud to walk into my ten year 15 reunion and be like 16 mary: [yeah 17 ivy: [<> see all you kipeople with kids have to 18 go home early [((laugh)) 19 mary: [i know (0.1) you should be proud of 20 yourselfyou’ve accomplished <> a lot (7) high school reunion 2 1 mary: so next year is the reunion 2 ivy: oh my goodness 3 mary: so i gotta start pumping some iron or something 4 ivy: yeah 5 (0.5) 6 mary: i need to be looking real good 7 ivy: ((laugh)) 8 mary: cause i can’t walk in there and be like (0.4) 9 <> oh my gawd [what happened to her 10 ivy: [((laugh)) 11 mary: she gained like a hundred pounds 12 (0.9) 13 ivy: yeah wewe definitely don’t want to be looking bad 14 mary: ((laugh)) 15 ivy: ((laugh)) in (6), ivy creates a quotation of future dialogue for herself for the high school reunion in lines 17-18 (see all you kipeople with kids have to go home early). while the situation may be exactly as she and mary are describing it (e.g. ivy may not have any children at the time she attends the reunion), it seems highly unlikely that ivy will remember to say this exact quote as she is watching the people who have kids at home leave the reunion celebration early. in the same way, mary creates future dialogue in (7) for her fellow classmates to say or think at the reunion in lines 9 and 11 (oh my gawd what happened to her; she gained like a hundred pounds); these words will most likely not be spoken or thought by anyone at the reunion when mary arrives. when speakers create this future dialogue, it is like they are using past experiences to propose hypothetical dialogue for the people involved in a future situation. however, not all future dialogue is based on a situation that will actually occur in the future. the next example is an interaction between ivy and jemma. they are 8 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 quotations in spoken discourse 9 discussing what they will say should anyone ask them what they did on super bowl sunday; this situation is one that will most likely not take place, so it only exists in the context of this conversation: (8) fake super bowl party 1 ivy: in bed at eight o’clock i’m like (0.7) <> ah i think i’ll read <> cause i’m 3 not really (0.2) too tired <> but 4 (0.4) 5 jemma: but if anyone else asks, you weren’t in bed until 6 like one 7 ivy: yeah exactly 8 jemma: because we had quite the [rager 9 ivy: [((laugh)) 10 jemma: at uh= 11 ivy: =bre[nt’s 12 jemma: [brent’s ((laugh)) 13 ivy: and i’ll be like <> i didn’t even watch like 14 thethe second half of the game (0.1) i don’t even 15 know who won ((laugh)) 16 jemma: because it was such a raging party? ((laugh)) 17 ivy: cause i was playing beer pong 18 jemma: oh yeah ((laugh)) should i give youyou knowa 19 couple of names to throw out? 20 ivy: ((laugh)) nah i don’t need names 21 jemma: i shouldi should sometime share some like random 22 stories and you can just adopt them as your own and 23 be like <> yeah and then this 24 happened 25 ivy: yeah definitely 26 jemma: ((laugh)) 27 ivy: ((h)) no i don’t want to start getting into too 28 detailed of a story [cause thenyeah 29 jemma: [(you’ll forget) later on and 30 you’ll be like <> oh crap 31 ivy: or then you knowsomeone will ask you about it and 32 you’ll be like uh <> yeah ((laugh)) 33 jemma: <> so what did ivy say exactly? [<> 34 um-hm that’s right, that’s what happened 35 ivy: [((laugh)) 36 i & j: ((laugh)) 37 ivy: so yeah nope (0.1) i’m just sticking to the story 38 we were at brent’s house (0.4) [and 39 jemma: [and we played beer 40 pong and won= 41 ivy: =it was yeahit was(0.3) it was like 42 cram packed and um (0.8) after like the third 43 quarter we got bored and went and played beer pong 44 and i don’t even know who won the football ga 45 super bowl ((laugh)) 46 jemma: <> sweet ((laugh)) wowi think we had a 47 pretty uh fun night there 48 ivy: hey i had a blast 49 (0.7) 50 jemma: sheesh (0.2) well i had a blast any[way 51 ivy: [well i did too 52 ((laugh)) 53 jemma: regardless of the fact that we weren’t at brent’s i 54 still had fun 9 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 10 55 ivy: that’s right this interaction has quite a bit of future dialogue that has been created by the two speakers, ivy and jemma. in lines 13-15, ivy creates speech for herself to tell anyone who asks that she did not even see the end of the super bowl game because she was having so much fun at the big party (that in reality she did not go to). in the same way, jemma creates speech for ivy to say about adopting jemma’s past stories as her own in lines 23-24. these examples, though, seem to act a lot like the examples we looked at in examples (6) and (7) about the high school reunion. ivy and jemma appear to be using past knowledge of how they would report being at a party to assume particular stances for similar future situations. the lines of dialogue that seem the most interesting and different from previous examples are 29-34: 29 jemma: [(you’ll forget) later on and 30 you’ll be like <> oh crap 31 ivy: or then you knowsomeone will ask you about it and 32 you’ll be like uh <> yeah ((laugh)) 33 jemma: <> so what did ivy say exactly? [<> 34 um-hm that’s right, that’s what happened just before this interaction, jemma proposed that she could give ivy stories to tell (previous stories from brent’s parties in the past), but ivy rejects that idea saying it would get too complicated the more details she wove into her story. that leads directly into the interaction above where ivy and jemma go back and forth using quotations to describe why having too many details for their super bowl “party” could be dangerous. it is interesting to note that in the interaction, jemma quotes fictive speech from the ‘voice’ of ivy in line 30 (oh crap), and ivy does the same thing by quoting fictive speech from the ‘voice’ of jemma in line 32 (uh yeah). these alone are noteworthy because they have already rejected the idea of adding details to the story, and yet both are quoting speech as if they did add details and then later forgot them. these quotations will never be spoken in this context simply because this context only exists in this conversation. lines 33-34 are rather interesting (but difficult to follow if taken out of context). jemma is doing her own future ‘voice’ and is quoting what she believes she would say if someone–who knew ivy and could have heard one version of the story from ivy–asked jemma what she did on the night in question. her response would be to ask “so what did ivy say exactly?” she would then wait for a response from the interrogator and then reply “um-hm. that’s right. that’s what happened.” again, this situation will never take place because a detailed story was rejected by the two girls earlier in the conversation. are these still demonstrations? they could be construed as demonstrations of hypothetical dialogue for participants involved in 10 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 quotations in spoken discourse 11 a fictive situation. even with this construal, though, it is becoming apparent that saying all quotations are demonstrations of something said, done, or thought is not an entirely true statement; it needs to be modified by saying that some quotations are demonstrations of something said, done in the past. other quotations can act as demonstrations of proposed hypothetical future dialogue or thought. myers (1999) labels this type of quotation as rhetorical reported speech and states that possible functions for this type of speech are to suggest counterarguments and to run a ‘thought experiment’ in which the participants can play out a rhetorical scenario using quoted material. in lines 29-34 of (8), both of these functions can be seen for the quotations used in the exchange. ivy and jemma are discussing why it would be a bad idea for ivy to randomly adopt stories of jemma’s to share with people, so these lines can be viewed as counterarguments to why they should not do this. also, because the idea of adopting stories had already been rejected, this can be viewed as a ‘thought experiment’ because ivy and jemma are using quotations to show what could happen should ivy decide to adopt jemma’s stories. 4. prosodic cues günther (1998) states that prosodic cues are used to achieve certain goals in the context of reported speech: “(i) to contextualize whether an utterance is anchored in the reporting world or the storyworld; (ii) to animate the quoted characters and to differentiate between the quoted characters; (iii) to signal the speech activities and the affective stance of the reported characters; and (iv) to comment on the reported speech as well as on the quoted characters” (21). while she has goals for the prosodic cues in reported speech, she does not specify whether these prosodic cues can be generalized to have meaning of their own (e.g. does breathy voice always signify reported inner speech?). do prosodic cues used for reporting inner speech and future dialogues have patterned uses? 4.1. inner speech in analyzing the data, it does not appear to be the case that a specific prosodic cue sets up inner speech; the prosodic cues depend on the mental state portrayed for the inner speech. for instance, the following reported inner speech taken from (1) is portraying a sense of being angry: 11 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 12 16 jemma: something (0.5) <> and i’m going <> 17 have you not remembered what [has happened 18 ivy: [((laugh)) 19 jemma: the past like (0.5) year and a half? 20 ivy: aw in this interaction, jemma is using her prosody to portray a sense of being upset with the other original participant in the story by making her voice higher and louder as she speaks. in the next example, though, ivy uses a breathy, high voice to portray a sense of possibly needing to do something but not necessarily wanting to do that thing right away: (9) calling r 1 ivy: ((h)) so yeah (0.3) (x) every once in a while i’ll 2 think about her and i’ll be like <> 3 oh i should probably call her ((laugh)) the two examples above both portray the use of a higher voice for the quoted inner speech; although jemma uses a louder voice to show her anger and ivy uses a breathy voice to talk about something she should do, the use of a higher voice is present in both. however, not all examples of inner speech use a higher voice pattern: (10) gym talk continued 1 ivy: ((laugh)) um so that’s a nice thing and (0.3) the 2 other thing is like, one they are so much cheaper 3 i only pay like 20 bucks a month 4 mary: oh that’s pretty [good 5 ivy: [where 24-hour fitness was making 6 methey were going to make me pay like 53 dollars 7 a month 8 mary: [<> what the heck? 9 ivy: [i’m like <> are you kidding? 10 mary: what are they smoking? 11 ivy: <> i know (0.1) [soand so 12 mary: [that’s not right 13 ivy: i get my classes that i want and (1.5) so yeah, i 14 really like it in the above example, ivy once again voices her inner thoughts through using the like quotative in line 9 (are you kidding) and the deictic pronoun you to refer to the gym and not her recipient; this time, though, her voice becomes louder for the quote to show her indignation at having to pay so much money for a gym membership. in this case, her voice does not become higher when she uses the quotation. as can be seen from the previous three examples, prosodic cues are different for reporting inner speech based on what type of mental state is being portrayed at the time. 12 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 quotations in spoken discourse 13 4.2. future speech as with reported inner speech, it appears that prosodic cues depend upon the proposed mental state for the future dialogue being conveyed by the speakers. in the following example, ivy does not change her voice pattern until the middle of the quoted material to portray a sense of wonder surrounding that particular quoted material: (11) grilled cheese 1 ivy: you’re going to be telling me that on my death bed 2 (0.1) your mom made the <> best grilled 3 cheese in lines 2-3, ivy accents the words best grilled cheese by changing her voice pattern to a breathy voice. from the earlier example about the fake super bowl party (8), different prosodic cues are used within the same interaction to convey reported future speech: 29 jemma: [(you’ll forget) later on and 30 you’ll be like <> oh crap 31 ivy: or then you knowsomeone will ask you about it and 32 you’ll be like uh <> yeah ((laugh)) 33 jemma: <> so what did ivy say exactly? [<> 34 um-hm that’s right, that’s what happened first, jemma uses a creaky, low-pitched voice in line 30 for the future quotation; but then ivy uses a high-pitched voice for a future quotation in line 32; right after that, jemma goes back to using her normal pitch but with a faster and then slower speed for the next chunk of reported future voice (lines 33-34). therefore, there do not appear to be links between the prosodic cues and the type of quoted material that follows; that information must be gleaned from contextual clues for the recipients to be able to comprehend the unfolding dialogue. instead of providing prosodic cues to signal what type of quoted speech will follow, these cues provide an insight into the speaker’s mental state or into what the speaker assumes is the mental state of a participant. as myers (1999) stated, “[w]hen the words are not said at all, it is apparently all the more important to convey how they would have been said” (587). therefore, speakers use prosodic cues as tools for conveying to their recipients how these words could or would have been said instead of as a cue to the onset of a quotation, which is also discussed in klewitz and couper-kuhlen (1999). 5. conclusion 13 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 14 much of the past research done on quotations in spoken discourse has focused on quotations that quote previously uttered material. however, many quotations in spoken discourse can quote material not spoken aloud in the past or material that is created within the context of a fictitious world. these quotations can still be said to be acting as demonstrations, but they appear to be acting as demonstrations of mental states rather than demonstrations of a particular situation. when using these special quotations, speakers do not appear to be employing specific prosodic cues to make the quote understood; rather, the recipient must rely on contextual cues to fully comprehend the use of the quotation. these quotations that are used to quote the unspoken act like any other quotation used in conversational speech with the use of appropriate deixis, quotatives, and prosody to convey affect; however, their function within conversational speech seems to be in a class of their own. references besnier, niko. 1993. “reported speech and affect on nukulaelae atoll.” in responsibility and evidence in oral discourse. jane h. hill and judith irvine (eds.), studies in the social and cultural foundations of language, 15: 161181. cambridge: cambridge university press. buttny, richard. 1997. “reported speech in talking race on campus.” human communication research 23: 477-506. clark, herbert h. and richard j. gerrig. 1990. “quotations as demonstrations.” language 66: 764-805. couper-kuhlen, elizabeth. 1998. “coherent voicing. on prosody in conversational reported speech.” interaction and linguistic structures (inlist) no. 1. günther, susanne. 1998. “polyphony and the ‘layering of voices’ in reported dialogues: an analysis of the use of prosodic devices in everyday reported speech.” interaction and linguistic structures (inlist) no. 3. holt, elizabeth. 2000. “reporting and reacting: concurrent responses to reported speech.” research on language and social interaction, 33 (4): 425454. holt, elizabeth. 1996. “reporting on talk: the use of direct reported speech in conversation.” research on language and social interaction 29 (3): 219-245. klewitz, gabriele and elizabeth couper-kuhlen. 1999. “quote unquote? the role of prosody in the contextualization of reported speech sequences.” interaction and linguistic structures (inlist) no. 12. myers, greg. 1999. “unspoken speech: hypothetical reported discourse and the rhetoric of everyday talk.” text 19(4): 571-590. 14 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 quotations in spoken discourse 15 niemelä, maarit. 2005. “voiced direct reported speech in conversational storytelling: sequential patterns of stance taking.” sky journal of linguistics 18: 197-221. szczepek, beatrice. 2001. “prosodic orientation in spoken interaction.” interaction and linguistic structures (inlist) no. 27. szczepek, beatrice. 2000. “formal aspects of collaborative productions in english conversation.” interaction and linguistic structures (inlist) no. 17. 15 sams: quoting the unspoken published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 16 appendix: key to transcriptions underline stressed/emphasized by speaker ? rising intonation . falling intonation break in sentence (normally used to revise what was just said) , micropause (smaller than 0.1 seconds) (0.2) length of pause in seconds [ overlapping speech = latched speech ((laugh)) non-speech sound <> voice quality (x) transcriber unable to hear the word spoken (word) transcriber’s best guess at what was said ((h)) outbreath 16 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/3 doi: https://doi.org/10.25810/7xd8-5581 colorado research in linguistics 6-2007 quoting the unspoken: an analysis of quotations in spoken discourse jessie sams recommended citation untitled a corpus study on the item-based nature of early grammar acquisition colorado research in linguistics volume 17 6-2004 a corpus study on the item-based nature of early grammar acquisition adam hodges carnegie mellon university valerie krugler stanford university deborah law university of colorado boulder follow this and additional works at: https://scholar.colorado.edu/cril part of the linguistics commons this working paper is brought to you for free and open access by linguistics at cu scholar. it has been accepted for inclusion in colorado research in linguistics by an authorized administrator of cu scholar. for more information, please contact cuscholaradmin@colorado.edu. recommended citation hodges, adam; krugler, valerie; and law, deborah (2004) "a corpus study on the item-based nature of early grammar acquisition," colorado research in linguistics: vol. 17. doi: https://doi.org/10.25810/ekag-3410 available at: https://scholar.colorado.edu/cril/vol17/iss1/2 https://scholar.colorado.edu/cril?utm_source=scholar.colorado.edu%2fcril%2fvol17%2fiss1%2f2&utm_medium=pdf&utm_campaign=pdfcoverpages https://scholar.colorado.edu/cril/vol17?utm_source=scholar.colorado.edu%2fcril%2fvol17%2fiss1%2f2&utm_medium=pdf&utm_campaign=pdfcoverpages https://scholar.colorado.edu/cril?utm_source=scholar.colorado.edu%2fcril%2fvol17%2fiss1%2f2&utm_medium=pdf&utm_campaign=pdfcoverpages http://network.bepress.com/hgg/discipline/371?utm_source=scholar.colorado.edu%2fcril%2fvol17%2fiss1%2f2&utm_medium=pdf&utm_campaign=pdfcoverpages https://scholar.colorado.edu/cril/vol17/iss1/2?utm_source=scholar.colorado.edu%2fcril%2fvol17%2fiss1%2f2&utm_medium=pdf&utm_campaign=pdfcoverpages mailto:cuscholaradmin@colorado.edu a corpus study on the item-based nature of early grammar acquisition adam hodges1, valerie krugler2, and deborah law1 1university of colorado, 2stanford university this paper explores the item-based nature of child language acquisition by examining data from the childes database (macwhinney 2000). two studies are explicated: the first uses pooled data from several children, and the second follows a single child longitudinally. the results show that the learning of the complex construction consisting of a main clause followed by an infinitival compliment, e.g. i want to play, center around a single verb, want, even though other candidate verbs exist in the children’s vocabulary. we provide empirical evidence to show that children initially learn grammar via itembased units and gradually break down complex constructions as units into smaller pieces in a process that leads towards the organization of language into the abstract categories consistent with a fully competent adult grammar. 1 introduction how do children acquire the grammar of their language? chomsky (1957) theorizes that a child is born with an innately equipped language module (cf. chomsky 1981, 1986a, 1986b). following chomsky, pinker posits a priori grammatical categories, such as verbs and nouns, in his proposals for semantic bootstrapping and linking rules (1984, 1989; cited in hoff 2001: 250, cf. pinker 1994). other nativists such as bower and wexler (1992, cited in tomasello 2000b: 160) posit a more gradual turning on of genes and syntactic structures. innatist proposals such as these base their starting point on the idea that the nature of children’s grammar is simply a miniature version of a fully competent adult. yet, from the outside perspective of an adult, child speech could be perceived as grammatical while operating quite differently during the developmental process of learning a first language. tomasello (1992) points out the highly item-based nature of child’s speech. rather than requiring the presence of innate categories or abstract structures, “children can also produce ‘grammatical’ language by simply reproducing the specific linguistic items and expressions (e.g. specific words and phrases) of adult speech, which are, by definition, grammatical” without productively using those items (tomasello 2000b: 156). for example, diessel and tomasello (2001) show that in early patterns of child usage, complement clauses are generally introduced by formulaic frames centered around only a few verbs, which they term constructional islands, or verb islands. their data demonstrate an early reliance on item-based patterns, with more novel uses occurring as experience with the language increases. thus, with regard to early syntactic development, it is important to ask whether children productively create utterances based on an abstract grammar or mainly reproduce item-based units as an initial step in the language learning process. if they possess an abstract grammar, they should be able to productively manipulate constituent categories, such as verbs. if they are merely reproducing constructions as units, then such productivity should be lacking in early stages and gradually increase over time. colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by adam hodges, valerie krugler, and deborah law. 1 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 this paper explores these ideas with data from the childes database (macwhinney 2000) in two separate studies to illustrate trends associated with a usagebased understanding of early syntactic development. we focus on variations of the complex construction consisting of a main clause followed by an infinitival complement, e.g. i want to play, he needs to eat, etc. the first study (section 2) uses pooled data from several children, while the second study (section 3) focuses on a longitudinal study of a single child. we provide evidence consistent with a usage-based understanding of child language development. that is, (1) the children in this study demonstrate early usage of this construction centered around a single verb, want, even though other verbs are present in their vocabulary. (2) the input language provides a preponderance of exemplars centered around the verb want but nevertheless shows a more level usage of different verbs in the main clause slot, a usage pattern that the children move toward in the course of development. (3) instances of this construction where the same subject is present in both the main clause and infinitival complement (i want to play versus i want him to play) are learned first; the ability to switch subjects from the main clause to infinitival clause comes much later in development, and when it does, it initially follows the same pattern demonstrated with single subject uses of the construction, relying on the verb want in the main clause before productively adding other verbs to that slot. a general discussion on the findings from both studies is provided in section 4. 2 initial study with pooled data from seven corpora the first study uses pooled data from several children and focuses specifically on first person usages of the main-clause-plus-infinitival-complement construction, e.g. i want to play, i need to eat, etc. we make two primary claims based on the results. (1) before 3 years of age, children rely on a single verb, want, even though other candidate verbs exist in their vocabulary, i.e. have, got, like, need. (2) after 3 years of age, children move away from their reliance on the single verb want in the main clause, and add a second verb have to the possibilities. this shift demonstrates increased productivity for this construction. 2.1 methodology the pooled corpus consists of data selected from the childes database (macwhinney 2000), and incorporates data from seven corpora, bates (1988; carlsonluden 1979), belfast (henry 1995; wilson and henry 1998), bloom (1970; bloom and lightbown 1974; bloom et al 1975), bloom (1973), bohannon (bohannon and marquis 1977; stine and bohannon 1983), clark (1978a, 1978b, 1979, 1982a, 1982b), and hall (hall et al 1984; hall et al 1981; hall and tirre 1979). all children in this combined corpus are between the ages of 1 year 1 month and 4 years 5 months. the corpus includes 304 data sessions from 67 children, a total of 86,010 child utterances. the original transcripts include varying degrees of additional notation with respect to nonspeech events, pauses, unclear utterances, etc. we preprocessed the data with a normalization script to remove most of these inconsistencies. to determine a frequently used frame to focus the study, sets of n-grams were generated for each age group, each age group represents approximately a six-month age 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 3 range. n-grams are collocations of terms within an utterance. bigrams (2 term collocates), trigrams, 4-grams, and 5-grams were generated to find frequent collocates in each group. one of the most frequent collocations across age groups above 2;0 is i want to, which upon further investigation generalized to i (don’t) v vp-inf (first person singular pronoun followed by an optional don’t, followed by a verb from a limited list, with an infinitive verb phrase complement.) this generalized frame was chosen as the focus for this study because of its relative frequency and complexity. regular expression search patterns were used to identify instances of this construction, 225 instances total. the regular expression patterns took into account another inconsistency in the data sample: the tendency for some transcriptions to contain more casual expressions of colloquial speech (e.g. ‘hafta’, ‘want ta’), while transcribed to a more formal style of speech (e.g. ‘have to’, ‘want to’.) the results of the regular expression searches were hand-checked to verify that all utterances considered aligned with the desired construction. eighteen utterances were discarded during the hand-checking. 2.2 results the construction: i (don’t) v vp-inf, and candidate verbs. in the complex construction, i (don’t) v vp-inf, the pooled data reveal a set of five verbs that are used in the main clause by the children: want, have, got, like, need. importantly, each of these verbs is found in the vocabulary of each age group studied, even though the frequencies vary markedly in their appearance in the construction studied. (n.b. no attested examples of this construction occur in the children below 2 years old. the complete set of data is given in appendix 1.) tables 1 and 2 show the frequency of occurrence of the five verbs found in this construction—want, have, got, like, need. as table 1 shows, want is used 84% of the time by the children in the 2;0-2;5 age group, 86% of the time by the 2;6-2;11 age group, 46% of the time by the 3;0-3;5 age group, 47% of the time by the 3;6-3;11 age group, and 54% of the time by the 4;0-4;5 age group. the use of the verb have is broken down as follows: 11% for the 2;0-2;5 age group, 4% for the 2;6-2;11 age group, 48% for the 3;0-3;5 age group, 47% for the 3;63;11 age group, and 42% for the 4;0-4;5 age group. the remaining verbs—got, like, need—make up a small percentage of use across all age groups, as shown in table 1. table 2 splits the numbers into two age groups: under 3 years old vs. 3 years old and above. we highlight this split since age 3 is an obvious dividing line in the children’s usage (as evidenced in table 1.) children below 3 years old use want 85% of the time in the main clause of this construction, while children 3 years old and above use want 48% of the time. in children 3 years old and above, the usage of have increases from 8% to 46%. the remaining verbs—got, like, need—make up a small remainder of the usage in both age groups. figures 1 through 4 illustrate the major trends from tables 1 and 2. figure 1 shows the usage of want across the age groups, as shown by the numbers in table 1. figure 2 illustrates the usage of want below vs. above 3 years, as detailed in table 2. figure 3 illustrates the trend that children over 3 years use verbs other than want more frequently in the main clause of this construction; and figure 4 shows the breakdown of the verbs used by children over 3 years. 3 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 table 1. distribution of the verbs used in the main clause by children across the age groups age 1;6-1;11 2;0-2;5 2;6-2;11 3;0-3;5 3;6-3;11 4;0+ verb # % # % # % # % # % # % want 0 0% 51 84% 60 86% 23 46% 8 47% 14 54% have 0 0% 7 11% 3 4% 24 48% 8 47% 11 42% got 0 0% 2 3% 6 9% 0 0% 0 0% 0 0% like 0 0% 1 2% 0 0% 2 4% 0 0% 1 0% need 0 0% 0 0% 1 1% 1 2% 1 6% 1 4% total 0 0% 61 100% 70 100% 49 100% 17 100% 27 100% table 2. distribution of the verbs used in the main clause by children below and above 3 years age under 3 years old 3 years old and above verb # % # % want 111 85% 45 48% have 10 8% 43 46% got 8 6% 0 0% like 1 1% 3 3% need 1 1% 3 3% total 131 100% 94 100% figure 1. use of ‘want’ in the main clause across age groups use of 'want' in main clause 84% 86% 46% 47% 54% 0% 20% 40% 60% 80% 100% age groups 2;0-2;5 2;6-2;11 3;0-3;5 3;6-3;11 4;0-4;5 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 5 figure 2. use of ‘want’ in the main clause by children below 3 years vs. children above 3 years use of 'want' in main clause 85% 48% 0% 20% 40% 60% 80% 100% age under 3 years over 3 years figure 3. use of verbs other than ‘want’ by children below 3 years vs. children above 3 years use of verbs other than 'want' 15% 52% 0% 10% 20% 30% 40% 50% 60% age under 3 years over 3 years figure 4. distribution of the verbs used in the main clause by children above 3 years old verbs used by children above 3 years 48% 46% 0% 3% 3% 0% 10% 20% 30% 40% 50% 60% verbs want have got like need 5 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 2.3 conclusions these data lead to the following conclusions. 1) before 3 years of age, children rely on a single verb, want, even though other candidate verbs exist in their vocabulary, i.e. have, got, like, need. 2) after 3 years of age, children move away from their reliance on the single verb want in the main clause, and add a second verb have to the possibilities. this shift demonstrates increased productivity for this construction. these conclusions suggest that children initially learn the main clause in this construction as a single unit. as their language develops, they gradually break down this entrenched unit into smaller pieces. past age three, they are able to apply other verbs to the construction, notably the verb have, in addition to want. this progression illustrates a gradual move toward the full competence characterized by adult grammar, a competence that demonstrates the ability to substitute constituents within larger clauses. 3 follow-up longitudinal study of a single child while our initial study uses pooled data and looks solely at utterances of child output, the purpose of our second study is to examine this particular construction again but in a slightly different light. (1) the construction we search for still involves a main clause followed by an infinitival compliment, but our search pattern captures usages of this construction across the paradigm rather than simply focusing on first person uses (i.e. i want to play and you want to play are both captured). (2) in addition to looking at uses of this construction where the subject is the same in both the main clause and infinitival clause, we carry out an additional search to find all uses of this construction where a noun is present between the main clause verb and infinitival marker, which demonstrates a switch in person between the subject of the main clause and subject of the infinitival clause. (3) we follow the development of a single child, rather than pooling data from several children, and also look at the adult input received by that child. by looking at variations of this construction and focusing on the development of a single child and the input language he receives, we provide another illustration of the item-based nature of the early stages of grammar acquisition. a single child was chosen for this study to avoid the problems of conflating the development of several children. an inherent problem in this type of approach, and indeed in any work done on child language development, involves the lack of sufficient longitudinal data. tomasello and colleagues are attempting to overcome this problem by collecting longitudinal data that represents a consistent and detailed look at the development of individual children. however, for the purposes of this study, we chose one of the best longitudinal corpora available on the childes database (macwhinney 2000), the corpus of adam collected by brown (1973). until sufficient longitudinal data from several children are available, we feel the pooled approach in section 2 and the individual approach in this section provide a starting point and evidence for a general trend worth further examination in the future. 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 7 we make the following claims based on the results of this section. (1) the child under study, adam, demonstrates the general pattern illustrated in section 2 where initial uses of the main-clause-plus-infinitival-complement construction are centered around the verb want even though other verbs are in his vocabulary. (2) the input language provides a preponderance of exemplars centered around the verb want but nevertheless shows a more level usage of different verbs in the main clause slot, a usage pattern that adam moves towards in his development. (3) adam’s usage of this construction relies heavily on same subject uses; the ability to switch subjects from the main clause to infinitival clause comes much later in development, and when it does, it initially follows the same pattern demonstrated with single subject uses of the construction, relying on the verb want in the main clause before productively adding other verbs to that slot. 3.1 methodology our data comes from a longitudinal study of the child adam conducted by brown (1973) and available as a part-of-speech tagged corpus on the childes database (macwhinney 2000). “adam was the child of a minister and an elementary school teacher. his family was middle class and well educated. though he was black, he was not a speaker of american black english, but of standard american. there are 55 files in the adam corpus and his age ranges from 2;3 to 4;10” (macwhinney 2000, based on brown 1973). the data indicate the age of adam at time of recording. the total number of utterances from adam in the corpus is 46,722, and the total number of utterances from caretakers is 114,081. while the documentation in the database states that data for adam ends at age 4;10, file number 52 lists the age of collection as 5;2. due to this discrepancy, we excluded file number 52 from the results. we used a perl script (appendix 2) to search the files (which are marked by age) for the presence of the following construction: (noun) (don’t) v vp-inf. regular expression search patterns were used to identity instances of this construction using the five potential candidate verbs identified in section 2—want, have, like, need, and got. the first four of these verbs were found in the speech of adam and his caretakers. the regular expression patterns took into account an inconsistency in the data— the tendency for some transcriptions to contain more casual expressions of colloquial speech (e.g. ‘hafta’, ‘want ta’), while transcribed to a more formal style of speech (e.g. ‘have to’, ‘want to’). the following is the regular expression used to search for the verb ‘want’ in its different forms: /^%mor:.*pro\|.*v\|wan(t|ts|ted|ting)( |~-)inf\|t(o|a) v\|/. each tagged line starts with %mor:, and each tag is in front of the word separated with a pipe. infinitives are separated from the main clause verb by ~-. counts of each verb occurring in the main clause were made for both the child and his caretakers. the perl script in appendix 2 includes the additional regular expressions used to search for the other candidate verbs in this construction. samples of the output data are provided in appendix 3. an additional search was made for a variation of this construction: (noun) (don’t) v noun vp-inf. that is, the same phrase but with a noun inserted between the main verb and infinitive marker to indicate a switch of subject from the main clause to the infinitival clause, e.g. i want him to go. the following is the regular expression used to search for the verb ‘want’ in this search: /^%mor:.*pro\|.*v\|wan(t|ts|ted|ting).*inf\|t(o|a) v\|/. 7 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 3.2 results the construction noun (don’t) v vp-inf, and candidate verbs. in the complex construction noun (don’t) v vp-inf, the adam corpus reveals four verbs that are used in the main clause by both adam and his caretakers: want, have, like, and need. importantly, each of these verbs is found in adam’s vocabulary throughout his development, even though the frequency of their appearance in the construction varies throughout that development. thus, it is not their lack in the lexicon that keeps them from appearing productively in this construction. table 3 shows the distribution of these verbs throughout adam’s development. he starts out relying on the verb want and gradually decreases this reliance as the construction begins to be used more productively with other verbs. from 2;3 to 2;11, adam uses want 95% of the time in this construction. from 3;0 to 3;11, the reliance on want decreases to 81% of the time as the second most used verb, like, increases to 13%. then from 4;0 to 4;10, want is used only 69% of the time, while have is used 18% of the time, like 10% of the time, and need 2% of the time in this construction. table 4 shows the corresponding input for this construction over the same time periods. the input language uses the verb want to a fairly large extent, providing the key exemplar that adam is picking up on; however, the other candidate verbs are also present to a large degree, providing the end result exemplar that adam’s development moves towards. table 3. distribution of the verbs used in the main clause of the target construction pro (don't) v vp-inf by adam across ages 2;3-2;11 3;0-3;11 4;0-4;10 # % # % # % want 56 95% 218 81% 175 69% have 1 2% 9 3% 46 18% like 2 3% 35 13% 26 10% need 0 0% 7 3% 5 2% total 59 100% 269 100% 252 100% 621 table 4. distribution of the verbs used in the main clause of the target construction pro (don't) v vp-inf from the input language across ages 2;3-2;11 3;0-3;11 4;0-4;10 # % # % # % want 77 63% 69 36% 43 39% have 21 17% 44 23% 34 31% like 21 17% 50 26% 30 27% need 4 3% 29 15% 4 4% total 123 100% 192 100% 111 100% 434 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 9 noun (don’t) v noun vp-inf. tables 5 and 6 show the usage of the more complicated form of the main-clause-plus-infinitival-complement construction. in this form of the construction, an additional noun is placed between the main clause verb and infinitival marker, which requires the speaker to productively manipulate the constituents that comprise the construction to indicate different subjects for the main and infinitival clauses. the data in table 5 show that adam does not have this ability before the age of 3 years, as no instances of this form of the construction occurred between 2;3 and 2;11. this is in contrast to the 59 instances of the simpler version of this construction that adam used during the same time period shown in table 1. as adam starts to use this more complicated form of the construction after 3 years, he shows a strong reliance on the verb want, in 86% and 87% of the occurrences from 3;0 to 3;11 and 4;0 to 4;10, respectively. table 6 shows the presence of this form of the construction in the input language across the same time periods. it is present at all stages of development, and provides a dominant number of exemplars using the verb want—91% during the age range of 2;3 to 2;11, 81% from 3;0 to 3;11, and 71% from 4;0 to 4;10. the presence of these additional uses of want adds to the number of exemplars of the general construction provided to the child during development. table 5. distribution of the verbs used in the main clause of the target construction pro (don't) v pro vp-inf by adam across ages 2;3-2;11 3;0-3;11 4;0-4;10 # % # % # % want 0 0% 24 86% 13 87% have 0 0% 1 4% 1 7% like 0 0% 2 7% 1 7% need 0 0% 1 4% 0 0% total 0 0% 28 100% 15 100% 47 table 6. distribution of the verbs used in the main clause of the target construction pro (don't) v pro vp-inf from the input language across ages 2;3-2;11 3;0-3;11 4;0-4;10 # % # % # % want 29 91% 21 81% 10 71% have 0 0% 3 12% 2 14% like 2 6% 2 8% 1 7% need 1 3% 0 0% 1 7% total 32 0% 26 100% 14 100% 73 9 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 3.3 conclusions the trends illustrated in these results lead to the following conclusions. 1) adam’s learning of this construction treats it as an item-based unit anchored around the verb want. as the child’s language skills develop, he is gradually able to treat the constituents of this complex construction as such and more productively use other verbs, and even switch subjects between the main clause and infinitival clause. 2) the high frequency occurrences of the verb want in the input language provide the most notable exemplars picked up on by the child. that is, the child begins learning the construction as a unit centered around the most frequently heard verb used in this construction by adults—want. these findings suggest that children initially treat this construction as a unit. as their language develops, they gradually break down this unit into smaller pieces. this progression illustrates a gradual move toward the full competence characterized by adult grammar, a competence that demonstrates the ability to productively manipulate various components within complex constructions. 4 general discussion a usage-based account of acquisition posits that language is built from the bottom up, from item-based units rather than innately present abstract categories. tomasello (2000b) states that “children imitatively learn concrete linguistic expressions from the language they hear around them, and then—using their general cognitive and socialcognitive skills—categorize, schematize and creatively combine these individually learned expressions and structures to reach adult linguistic competence” (tomasello 2000b: 156). the findings in this paper provide further evidence for this type of acquisition model. 4.1 the learning of language via item-based units the importance of larger-than-word units in language is an important idea stressed in cognitive linguistics. langacker (1987) explains, “with repeated use, a novel structure becomes progressively entrenched, to the point of becoming a unit” in grammatical organization (langacker 1987: 59). that is, language is composed of a ‘structured inventory of constructions.’ with this idea in mind, we posit that children are apt to learn language by pulling in such units into their own usage. therefore, their first attempts at producing their own utterances rely heavily on these constructions as units (see langacker 1987: 57 for a detailed discussion of units; also, see langacker 1987: 50 for a detailed discussion of entrenchment in grammar, and tomasello 2000a: 72 for its relation to language acquisition; cf. langacker 1991, 1998). the data from this paper show that children initially rely on exemplars centered around the verb want when learning the complex construction that involves a main-clause-plus-infinitivalcomplement. 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 11 this finding is backed up by previous research. for example, peters (1986) provides the examples of iwant and idontknow (cited in hoff, 2001: 210) as multiword phrases memorized as units in the lexicon of children. hoff (2001: 210) refers to these as “rote-learned wholes” and “jargon + word combinations.” lieven et al (1997) show that initial multi-word utterances of children are based on specific lexical patterns (cited in tomasello 2000: 157). and bolinger (1976) suggests that language learners, whether children or adults, often learn via collocations (cited in johnson, 2000: 193). bolinger (1977) presents examples of “collocation mixing” to show the process of internal analysis that then takes place in children as they gradually break up larger phrases into words (cited in johnson 2000: 192). similar to the case of the construction presented here, wh-word development has been shown to have formulaic beginnings. johnson (2000: 187) cites bellugi (1965), brown (1968, 1973), klima and bellugi (1966), and johnson (2000), as evidence for such a process. once a wh-frame is rote-learned, then the child abstracts the wh-word to new situations in what johnson’s (2000) data show is a gradual process. the end-result is competence in using the wh-words themselves in appropriate, novel situations. and in the case of negation, hoff (2001: 219) cites evidence that children often start out using whole forms such as can’t and don’t as negative markers before analyzing and breaking out the conflated auxiliaries (can, do) and negative marker (not). the data presented in this paper clearly show a similar pattern. the children produce the complex construction studied by relying on the verb want, a verb that is most frequently used in the adult input, as attested in the study of adam. rather than needing to master the complex grammatical abstractions involved in such a construction, the child can simply reproduce the most typical exemplar. once this complex construction is learned as an item-based unit, then the child begins to gradually abstract out the finer, individual components, and use the construction more productively—first adding other verbs to the main clause verb slot and then switching subjects between the main clause and infinitival clause. 4.2 from item-based units to abstract grammar the data in this study show a shift from a reliance on the use of the single verb want in the main clause toward an ability to more productively substitute other verbs into the appropriate slot. this progression is demonstrated in the results of both section 2 and section 3. section 2 shows that for the pooled data, there is a marked shift in productive ability with this construction after 3 years of age. the study of adam in section 3 demonstrates the same progression toward productivity over his development. clearly, the abstract category of verb is not being productively applied in the early usages demonstrated in the data, suggesting that such an innate category is lacking. in explaining his verb island hypothesis, tomasello (2000b: 157) claims that children have “lexically specific syntactic categories” rather than abstract syntactic categories. this is similarly demonstrated in our data, which point to the lexically specific nature of the construction studied in early grammatical development. the near exclusive reliance on the collocation of want with the surrounding constituents of the construction, suggests that children do not yet have a more abstract schema that treats this construction as being composed of categories (whether present or underlying) such as noun + verb + (noun) + inf-marker + verb. out of the five verbs used by 11 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 children in the initial study and the four verbs used by adam in this type of construction, all are present in the vocabulary during each age range examined. however, the children fail to substitute these plausible candidates into the main clause in this construction at earlier ages. productivity gradually increases as they get older and develop more experience with language in general and this construction in particular. children in the initial study begin to move away from the reliance on the verb want after age 3. however, they do not immediately exhibit free use of any of the candidate verbs, but rather move from reliance on one verb to reliance on two verbs— want and have. similarly, from 3;0 to 3;11, adam begins to replace his reliance on the single verb want by adding another verb to his repertoire—like. from 4;0 to 4;10, the verb have begins to play more of a role in adam’s productive ability. this demonstrates that productivity increases gradually, and the movement towards increased productivity involves a step-by-step process. this process can be thought of as one that involves learning separate constructions as units—first a construction with the verb want, then a construction with the verb like, then a construction with the verb have, and so on. rather than a complete, all-or-nothing shift towards knowledge of abstract categories, the child gradually builds up a repertoire of constructions until the point is reached when generalizations can be made about this set as a whole. then the child learns to abstract out the individual constituents and manipulate those constituents more productively. the end result is the fully productive, abstract grammar of an adult. the data presented in this study are similar to other findings. tomasello (2000b: 157) cites cross-linguistic studies which imply that the most frequent verb forms in the input language are first learned as unanalyzed items by children; then the inflections are gradually abstracted out, applied to other verbs and eventually abstracted to the rest of the forms in the verb paradigm (pizutto and caselli 1992, 1994; rubino and pine 1998; berman and armon-lotem 1995; berman 1982; macwhinney 1978; behrens 1998; allen 1996; gathecole 1999; stoll 1998). experimental studies cited by tomasello (2000b: 158-159) demonstrate this as well. studies by akhtar (1999), akhtar and tomasello (1997), brooks and tomasello (1999), berman (1993), dodson and tomasello (1998), maratsos et al (1987), olguin and tomasello (1993), pinker et al (1987), tomasello and brooks (1998) imply a gradual progression in children’s ability to extend novel verbs to constructions different from those in which the verbs were learned, such that children younger than 3 lack the productivity that children older than 3 are able to demonstrate. furthermore, samples of children’s overgeneralizations as compiled by bowerman (1982, 1988) and pinker (1989) predominantly occur after age 3, presumably once more abstract categories have been extracted from item-based units (cited in tomasello 2000b: 158). as tomasello states, “…it is clear that young children are productive with their early language in only limited ways. they begin by learning to use specific pieces of language and only gradually create more abstract linguistic categories and schemas” (tomasello 2000b: 159). finally, the exemplars provided in the input language are an important component in the process of learning grammar. the most frequently heard exemplars of the complex construction, as attested in the adam study, are centered around the verb want. the child picks up on this frequently heard exemplar and bases early uses of the construction on this verb. the frequency of the verb want in the input data is even greater when the more complex form of the construction that involves a switch of subject between the main 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 13 clause and infinitival clause is examined (section 3). while a young child is unable to pick out every constituent of that construction, he hears want enough so that his early production of the simpler version of the construction is based upon that ‘verb island,’ or ‘constructional island’ (diessel and tomasello 2001, cf. elsewhere). 5 summary this study provides empirical evidence, pertaining to the complex construction comprised of a main clause plus infinitival complement, for the idea that children initially learn grammar via item-based units, and then gradually break down complex constructions as units into smaller pieces in a process that leads towards the organization of language into abstract categories consistent with a fully competent adult grammar. this is a pattern “consistent with a more constructivist or usage-based model in which young children begin language acquisition by imitatively learning linguistic items directly from adult language, only later discerning the kinds of patterns that enable them to construct more abstract linguistic categories and schemas” (tomasello 2000b: 158). the evidence provided here creates a difficult obstacle for a high-innateness model of language acquisition, which relies on a priori abstract categories in the mind of the young language learner. if these categories were present from the beginning, then they would be expected to result in much greater productivity than what these data show. 13 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 14 references akhtar, n. and m. tomasello. 1997. “young children’s productivity with word order and verb morphology.” developmental psychology, 33: 952-965. akhtar, n. 1999. “acquiring basic word order: evidence for data-driven learning of syntactic structure.” journal of child language, 26: 261-278. allen, s. 1996. aspects of argument structure acquisition in inukitut, john benjamins. bates, e. and i. bretherton and l. snyder. 1988. from first words to grammar: individual differences and dissociable mechanisms. cambridge, ma: cambridge university press. behrens, h. 1998. “how difficult are complex verbs? evidence from german, dutch, and english.” linguistics, 36: 679-712. bellugi, u. 1965. “the development of interrogative structures in children’s speech.” in k. riegel, ed., the development of language functions, report no. 8, 103-137. ann arbor, mi: language development program. berman, r.a. and s. armon-lotem. 1995. “how grammatical are early verbs?” annales littéraires de l’université de franche-comté, 631 : 17-56. berman, r. 1982. “verb-pattern alternation: the interface of morphology, syntax, and semantics in hebrew child language.” journal of child language, 9: 169-191. —. 1993. “marking verb transitivity in hebrew-speaking children.” journal of child language, 20: 641-670. bloom, l. and l. hood and p. lightbown. 1974. “imitation in language development: if, when and why.” cognitive psychology, 6, 380–420. bloom, l. and p. lightbown and l. hood. 1975. “structure and variation in child language.” monofigures of the society for research in child development, 40, (serial no. 160.) bloom, l. 1970. language development: form and function in emerging grammars. cambridge, ma: mit press. —. 1973. one word at a time: the use of single-word utterances before syntax. the hague: mouton. bohannon, j.n. and a.l. marquis. 1977. “children’s control of adult speech.” child development, 48, 1002–1008. bolinger, dwight. 1976. “meaning and memory.” forum linguisticum, 1, 1-14. —. 1977. “idioms have relations”. forum linguisticum, 2, 157-169. bower, h. and wexler, k. 1992. “bi-unique relations and the maturation of grammatical principles.” natural language linguistic theory, 10: 147-187. bowerman, m. 1982. “reorganizational processes in lexical and syntactic development.” in l. gleitman and e. wanner, eds., language acquisition: the state of the art, 319-346. cambridge, england: cambridge university press. —. 1988. “the ‘no negative evidence’ problem: how do children avoid constructing an overgeneral grammar?” in j.a. hawkins, ed., explaining language universals, 73101. basil blackwell. brooks, p. and m. tomasello. 1999. “young children learn to produce passives with nonce verbs.” developmental psychology, 35: 29-44. 14 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 15 brown, roger. 1968. “the development of wh questions in child speech.” journal of verbal learning and verbal behavior, 7, 279-290. —. 1973. a first language: the early stages. cambridge, ma: harvard university press. carlson-luden, v. 1979. causal understanding in the 10-month-old. unpublished doctoral dissertation. university of colorado at boulder. chomsky, noam. 1957. syntactic structures. the hague: mouton. —. 1981. lectures on government and binding. dordrecht: foris. —. 1986a. knowledge of language: its nature, origin and use. new york: praeger. —. 1986b. barriers. cambridge: mit press. clark, e.v. 1978a. “awareness of language: some evidence from what children say and do.” in r. j. a. sinclair & w. levelt, eds., the child’s conception of language. berlin: springer verlag. —. 1978b. “discovering what words can do.” in w. j. d. farkas & k. todrys, eds., papers from the parasession on the lexicon. chicago: chicago linguistic society. —. 1979. “building a vocabulary: words for objects, actions and relations.” in p. fletcher & m. garman, eds., language acquisition: studies in first language development. new york: cambridge university press. —. 1982a. “language change during language acquisition.” in m. e. lamb & a. l. brown, eds., advances in child development: vol. 2. hillsdale, nj: lawrence erlbaum associates. —. 1982b. “the young word maker: a case study of innovation in the child’s lexicon.” in e. wanner & l. r. gleitman, eds., language acquisition: the state of the art. cambridge, ma: cambridge university press. diessel, holger and michael tomasello. 2001. “the acquisition of finite complement clauses in english: a corpus-based analysis.” cognitive linguistics, 122: 97-141. dodson, k. and m. tomasello. 1998. “acquiring the transitive construction in english: the role of animacy and pronouns.” journal of child language, xx: xxx-. gathecole, v. et al. 1999. “the early acquisition of spanish verbal morphology : across-the-board or piecemeal knowledge ?” international journal of bilingualism, 3: 138-182. hall, w.s. and w.c. tirre. 1979. the communicative environment of young children: social class, ethnic and situational differences. champaign, il: university of illinois. hall, w.s. and w.e. nagy and g. nottenburg. 1981. situational variation in the use of in-ternal state words. champaign, il: university of illinois. hall, w.s. and w.e. nagy and r. linn. 1984. spoken words: effects of situation and social group on oral word usage and frequency. hillsdale, nj: erlbaum. henry, a. 1995. belfast english and standard english: dialect variation and parameter setting. new york: oxford university press. hoff, erika. 2001. language development, belmont, ca: wadsworth/thomson learning. johnson, carolyn e. 2000. “what you see is what you get: the importance of transcription for interpreting children’s morphosyntactic development.” in l. menn and n.b. ratner, eds., methods for studying language production. mahwah, nj: lawrence erlbaum associates, inc. 15 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 16 klima, e. and u. bellugi. 1966. “syntactic regularities in the speech of children.” in j. lyons and r. wales, eds., psycholinguistics papers, 183-208. edinburgh: university press. (reprinted in a. bar-adon and w. leopold, eds., child language: a book of readings, 412-424, englewood cliffs, nj: prentice hall.) langacker, ronald. 1987. foundations of cognitive grammar: volume i. stanford: stanford university press. —. 1991. foundations in cognitive grammar: volume 2. stanford: stanford university press. —. 1998. “conceptualization, symbolization, and grammar.” in michael tomasello, ed., the new psychology of language: cognitive and functional approaches to language structure. mahwah, nj: lawrence erlbaum associates, inc. lieven, e. 1997. “variation in a cross-linguistic context.” in d. slobin, ed., the crosslinguistic study of language acquisition, vol. 5: 233-287. erlbaum. macwhinney, brian. 1978. “the acquisition of morphophonology.” monogr. soc. res. child dev., no. 43. —. 2000. the childes project: tools for analyzing talk. 3rd edition. vol. 2: the database. mahwah, nj: lawrence erlbaum associates. maratsos, m. et al. 1987. “a study in novel word learning: the productivity of the causative.” in b. macwhinney, ed., mechanisms of language acquisition, 89-114, erlbaum. olguin, r. and m. tomasello. 1993. “twenty-five month old children do not have a grammatical category of verb.” cognitive development, 8: 245-272. peters, anne. 1986. “early syntax.” in p. fletcher and m. garman, eds., language acquisition, 2nd edition: 307-325. cambridge, england: cambridge university press. pinker, steven et al. 1987. “productivity and constraints in the acquisition of the passive.” cognition, 26: 195-267. pinker, steven. 1984. language learnability and language development. cambridge, ma: harvard university press. —. 1989. learnability and cognition: the acquisition of argument structure. cambridge, ma: mit press. —. 1994. the language instinct: how the mind creates language. new york: morrow. pizutto, e. and c. caselli. 1992. “the acquisition of italian morphology.” journal of child language, 19: 491-557. —. 1994. “the acquisition of italian verb morphology in a cross-linguistic perspective.” in y. levy, ed., other children, other languages: 137-188. erlbaum. rubin, r. and j. pine. 1998. “subject-verb agreement in brazilian portuguese: what low error rates hide.” journal of child language, 25: 35-60. stine, e.l. and j.n. bohannon iii. 1983. “imitations, interactions, and language acquisition.” journal of child language, 10, 589–603. stoll, s. 1998. “the acquisition of russian aspect.” first language, 18: 351-378. tomasello, michael. 1992. first verbs: a case study in early grammatical development, cambridge, england: cambridge university press. —. 2000a. “first steps toward a usage-based theory of language acquisition.” cognitive linguistics, 11-1/2: 61-82. —. 2000b. “the item-based nature of children’s early syntactic development.” trends in cognitive science, vol. 4, no. 4: 156-163. 16 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 17 wilson, j. and a. henry. 1998. “parameter setting within a socially realistic linguistics.” language in society, 27, 1–21 17 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 18 appendix 1: data from initial study all instances of the i (don’t) v vp-inf construction used by children in this study 2;0 would you # i have to put the pen back in my pocketbook. 2;1 [?] i have to put um +... 2;1 i i don't want to go anywhere. 2;1 mummy mummy i wanna [: want to] help daddy cause i'm a good girl. 2;1 and i want to come too. 2;1 and i want to play wuh wi with [?] some 2;1 but i have to put a bench for him. 2;1 dat's i wanna p i wanna i wanna hang this up wif dis. 2;1 duh floor is dirty # so i have to um bwush it wif my 2;1 get a screwdriver # i have ta find one. 2;1 i i wanted to put him in in. 2;1 i wanna be dey cook i wanna i wanna cook dem. 2;1 i wanted to uhh i want to +... 2;1 n+no # i wanna give um no # i'll go do it. 2;1 no # af no no first i have to invite all my 2;1 no # i wanna play play+dough. 2;1 no # i want to put him next to +... 2;1 no no # i wanna put some i wanna put some juice in dem 2;1 no no no # i wanna give duh froggie dat banana. 2;1 nope # i got ta get another one. 2;1 somebody can sit # um [?] oh # i have to wear uh hat. 2;1 this i wanna i don't wanna see mister wogers. 2;1 with these tools # i got ta fix em. 2;1 yeah # uh um i wanna go play now. 2;1 yeah i like beer cos i like to xxx. 2;2 i i wanna xxx xxx in the water. 2;2 no # i wanna go uh see uh baby. 2;2 this my writing # i want ta write. 2;2 uh mommy and i i wanna put on uh that. 2;3 i wanna go but uh and see it uh turn around that 2;3 i wanna eat it and like that # i wanna broken and all page 2;3 aw # i i wanna fo. 2;3 he's doing work in the garage # i wanna go walk up and see the xxx 2;3 i think i wanna go bed. 2;3 i think i wanna go binky bed. 18 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 19 2;3 i want i wanna bite some. 2;3 i wanna i want i wanna i wanna's 2;3 i wanna play+dough # i wanna play+dough in the room. 2;3 i want it # i wanna play+dough. 2;3 tape # i wanna play tape that go in a box. 2;3 xxx xxx what in a box &i i wanna climbing &i &i in mommy's room? 2;3 yeah # i wanna carry. 2;3 yeah # i wanna put i wanna that record off. 2;4 to do that # because i wanna pah pah pah pah! 2;4 i wanna get the book. 2;4 i i wanna punch out this. 2;4 i i want to write ana's name # i want to write ana's 2;4 kay # i want to get +... 2;4 yeah # xxx i wanna big kitty. 2;5 because i wanna go in the water # daddy # i wanna go out. 2;5 because i wanna get down. 2;5 because the wind blow down there to the grass because i wanna go 2;5 i i i wanna put dis on dat fing # and 2;5 i want do # i want to wun on duh stickers now. 2;5 i want # i want ta play airplane in the sky with mine # okay? 2;5 i want to play with daddy machine # i want to i wanna 2;5 in the sky because i wanna get down. 2;5 no # i i wanna wear pants and uh short # okay? 2;5 no # i wanna pour some more. 2;5 yes # i wanna go play. 2;6 i want ta write with it. 2;6 and so i have to fix it. 2;6 i want to wun awound. 2;6 i want ta put this mommy on the +... 2;6 i want to play them. 2;6 i want to go outside. 2;6 i want ta give i want ta give you a piece # a # paper # of 2;6 help me walk i jump by myself # i wanna jump on 2;6 i want see how duh clock work # i wanna have 2;6 it broke # i have to fix it with other. 2;6 no # i wanna go in in in in 2;6 no # i wanna go wif mommy and daddy. 2;6 no # i want to go on dere. 19 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 20 2;6 no i xxx # i want ta be quiet in the bus. 2;6 xxx # i want ta write. 2;7 i want ta put this -. 2;7 and i wanna # i wanna fings &da you go 2;7 i want to go in duh store and buy some string 2;7 i wanna put it 2;7 want want i wanna get 2;7 ah # i want to go and see. 2;7 dere # i i wanna put it outside. 2;7 i i i want ta put it in the sky # now. 2;7 i wanna i wanna put it together # go 2;7 no # i got ta walk in here. 2;7 no i need to put my lipstick on first. 2;7 no i want ta write on your paper # ok. 2;7 no i wanna go to minky+bed # dat means i want to go to house. 2;7 now # i want to write. 2;7 now i can go up there # i wanna climb up here. 2;7 oh # i i have to wing duh phone first # i have to wing 2;7 oh ## better hurry # i got ta get some food. 2;7 ok # i want to write. 2;7 what's that i wanna [: want to] see # penguins show me # see 2;7 yeah # see if dis broke # i wanna put some words too 2;7 yeah ## i got ta get the doctor. 2;8 cos i wanna [: want to] put it in here. 2;8 i want ta wear that 2;8 but i but i said i wanna go to your house. 2;8 but i wanna i i wanna read it back here. 2;8 but i want to # uhh+oh # it fits. 2;8 i i wanna +... 2;8 i i wanna put it right dere next to your yips. 2;8 i wan i wanna jump on your waterbed. 2;8 i wan i wanna put it right here. 2;8 i want i wanna put it near here. 2;8 i want i wanna xxx in dere. 2;8 i wanna play wif i wanna play wif i wanna play wif dis. 2;8 like that # i got ta go there and can't fit. 2;8 momma # i want to go [?] up here. 2;8 no # i wanna put it on here. 20 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 21 2;8 uhh # i wanna go to room again. 2;8 uhh # i wanna put it in the house. 2;8 uhh i i wanna put it in the cup. 2;8 uhh uhh uhh i want to watch. 2;8 we i wanna play uh wecord in my room. 2;8 xxx i want to pick it. 2;8 yeah # i want ta wear you ring. 2;9 and also i wanna play dat fing. 2;9 and here's a book dat i want to to make. 2;9 and i can take it off because i wanna put it on here. 2;9 because i wanna um leave it like &da in case dis goes like 2;9 but i wanna put another one on &i and two fings. 2;9 don't # i i want i wanna you put it down. 2;9 i want duh duh ladder and i wanna show i wanna get dose # 2;9 now # i wanna fing wight here # okay ? 2;9 there's a car in there # i got ta go see it. 2;9 this was broken and i got ta fix it. 2;9 where when i wanna get off i just can jump off and then +... 3;0 [?] first # i hafta get uh game. 3;0 [?] i wanna play some playdough on some paper. 3;0 no i wanna weed to you. 3;0 and this is the bus because i like to go to the grandma's house. 3;0 another children could s i wanna put duh bed in +... 3;0 but i don't wanna do i don't wanna go in our my car! 3;0 but i wanna play wif duh families. 3;0 first i hafta get um this xxx this out # and this. 3;0 fo i wanna put [?] back on dere. 3;0 i think i wanna get this. 3;0 i wanna put it in i wanna put this string in there. 3;0 no # i i have to lie dem down [?] second. 3;0 no # i'll put it on i want to put it here. 3;0 now i have to put dem back in. 3;0 um but first i hadded to get like dis +... 3;0 um i have to uhh i wanna play wif uh little game. 3;0 yeah # i i wanna go pee. 3;1 because i haf to go to another place. 3;1 i i i oh i have to look for the sun # sun # nope [: no]. 3;1 i want to # i want to come! 21 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 22 3;1 because i wanna [?] s'm some more potato chips. 3;1 but but now i hafta put the legs. 3;1 cause i had them # i had to knock them down. 3;1 em # i xxx i had to xxx them down. 3;1 ha ha ha # xxx every day # and i have to &s uh # do all my homework 3;1 hey # i hafta get some chairs. 3;1 hey # i have to c take him somewhere. 3;1 hey # i wanna leave it closed. 3;1 i think i hafta pound it harder and harder. 3;1 i think i have to give this to amanda uhh uhh 3;1 i wan i wanna get dese off. 3;1 i wanna play i wanna play with something else. 3;1 mummy i want to play! 3;1 now i have to dry this up and i hafta go to another place. 3;1 oh # no # i haf i have to go on dis motor+cycle. 3;1 oh # no # oh # no # i hafta put dis in. 3;1 uh i have to keep them down here. 3;1 uhh no # dese are i wanna cover up duh elephant because 3;1 um # i want to be the fire brigade. 3;1 um i think i haf to wait. 3;1 well i want to go. 3;1 yeah # i hafta climb up here. 3;1 yeah # i hafta poke it uh little more. 3;1 yep # i have to go on dis motor+cycle # because we need to drive s 3;2 no # i like to get um uhh no # i have popcorn # but 3;2 yeah # and i wanna i wanna see it make duh sound. 3;5 i was standing # i don't need to look at the # these animals. 3;5 and then i want to go to my mummy. 3;5 get him back on my # on your knee # i want to xx. 3;6 i have to wash my hands but i need to go to the toilet as well. 3;6 a wee thing but i have to put all of them in. 3;6 and i want to put makeup on to me. 3;6 and you keep it closed # cause i have to get my bike with me. 3;6 get sand in # i have to get a new bucket and get sand in # to the 3;6 now i have to catch another one. 3;6 yeah but i have to get your breakfast ready. 3;7 i don't want one on there # i want to go on the floor. 3;7 no i don't want to catch a fish. 22 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 23 3;7 the wee ones aren't # where i had to keep them. 3;7 xxxx i # i don't want to play with xxx. 3;9 i will feed him # i # i want to bring my xx back xx xx. 3;9 mummy i want to keep my tracker coat on. 3;9 and i want to play with that # xx lift him again. 3;9 can i put it back in # i want to put it back in. 3;9 oh # i have to get this in! 4;0 i know what i want to play with xxx. 4;0 i'm going to be # i'm go # i want to put all these back in cos i'm 4;0 mummy i don't wanna [: want to] go on holidays. 4;0 and i have to get a trailor. 4;0 can't find what i have to put on. 4;0 cause i wanna [: want to] do #. 4;0 now if you open your bags and i wanna [: want to] see what's in 4;0 oh i # i have to close this up i need everything in it. 4;0 well # that's all i have to give him. 4;0 xxx i have to xxx. 4;0 xxx i wanna [: want to] go out and play xxx. 4;0 yep i have to get a new one. 4;1 so i have to be uh # this one. 4;2 do i have to catch anyone i like to? 4;2 no i want to watch a video with david. 4;2 yes and i want to play that. 4;4 i like it and i want to keep it. 4;4 cause i like # cause i like calling it neil # i want to call it 4;4 cause i wanted to build a city so i breaked it. 4;4 cause the man couldn't get down the stairs # i have to lift him. 4;4 no cause # cause i don't want to # to do it the day [: today]. 4;4 um # i wanna [: want to] have a look xxx. 4;4 well i need to build you a wee boat. 4;4 why do i have to go to granny kenny's? 4;4 yip # if the two dies i have to get # i have to call him grape and 4;5 well i don't want to go. 23 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 24 appendix 2: study 2 perl script to gather data for the search of (noun) (don’t) v vp-inf #!/usr/local/bin/perl use strict "vars"; my $filecount=55; my $name='adam'; my $childid='chi'; my $currentline; my $previousline=undef; my %childlines; my %otherslines; my %totalverbs; my %verbcounts; my $child='child'; my $others='others'; open(result,">>results$name.txt")||die "couldn't open : $!\n"; print result "child: $name\n"; print result "search pattern: pro (don't) v vp-inf\n"; for(my $i=1;$i<$filecount+1;$i++) { open(in,"$i$name.txt")||die "couldn't open <$i$name.txt>: $!\n"; %childlines=(); %otherslines=(); %totalverbs=(); while($currentline=) { chomp($currentline); findandprintage($currentline); matchregex($currentline,$previousline,$childid); totalverbcounts($currentline,$previousline,$childid); $previousline=$currentline; } %verbcounts=(); counteachverb(%childlines); printverbcount($child,%verbcounts); %verbcounts=(); counteachverb(%otherslines); printverbcount($others,%verbcounts); printdata($child,%childlines); printdata($others,%otherslines); printverbtotal($child,%totalverbs); } sub findandprintage { my ($line)=@_; if($line=~/^\@id:.*\|(\d;\d.(\d|))/) { print result "\n\n........................................\n\n"; print result "age: $1\n"; } } sub matchregex { 24 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 25 my ($current,$previous,$id)=@_; if($current=~/^%mor:.*v\|wan(t|ts|ted) inf\|t(o|a) v\|/) { if($previous=~/^*$id/) { $childlines{$previous}='want'; } else { $otherslines{$previous}='want'; } } elsif($current=~/^%mor:.*v\|ha(ve|s|d) inf\|t(o|a) v\|/) { if($previous=~/^*$id/) { $childlines{$previous}='have'; } else { $otherslines{$previous}='have'; } } elsif($current=~/^%mor:.*v\|g(et|ets|ot) inf\|t(o|a) v\|/) { if($previous=~/^*$id/) { $childlines{$previous}='get'; } else { $otherslines{$previous}='get'; } } elsif($current=~/^%mor:.*v\|nee(d|ds|ded) inf\|t(o|a) v\|/) { if($previous=~/^*$id/) { $childlines{$previous}='need'; } else { $otherslines{$previous}='need'; } } elsif($current=~/^%mor:.*v\|lik(e|es|ed) inf\|t(o|a) v\|/) { if($previous=~/^*$id/) { $childlines{$previous}='like'; } else { $otherslines{$previous}='like'; } } 25 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 26 } sub totalverbcounts { my ($current,$previous,$id)=@_; if($current=~/v\|wan(t|ts|ted)/) { if($previous=~/^*$id/) { $totalverbs{want}++; } } if($current=~/v\|ha(ve|s|d)/) { if($previous=~/^*$id/) { $totalverbs{have}++; } } if($current=~/v\|g(et|ets|ot)/) { if($previous=~/^*$id/) { $totalverbs{get}++; } } if($current=~/v\|nee(d|ds|ded)/) { if($previous=~/^*$id/) { $totalverbs{need}++; } } if($current=~/v\|lik(e|es|ed)/) { if($previous=~/^*$id/) { $totalverbs{like}++; } } } sub counteachverb { my $value; my (%verbhash)=@_; my @values=sort(values(%verbhash)); foreach $value (@values) { if($value=~/want/) { $verbcounts{want}++; } elsif($value=~/have/) { $verbcounts{have}++; } elsif($value=~/get/) 26 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 27 { $verbcounts{get}++; } elsif($value=~/need/) { $verbcounts{need}++; } elsif($value=~/like/) { $verbcounts{like}++; } } } sub printverbcount { my $element; my($person,%verb)=@_; print result "\n$person\n\n"; my @array=sort(keys(%verb)); foreach $element (@array) { print result "\t$element: $verb{$element}\n"; } } sub printdata { my $key; my ($person,%hash)=@_; my @total=sort(keys(%hash)); print result "\n$person data\n\n"; foreach $key (@total) { print result "\t$key\n"; } } sub printverbtotal { my $element; my($person,%totalverb)=@_; print result "\ntotal verb count for $person\n\n"; my @array=sort(keys(%totalverb)); foreach $element (@array) { print result "\t$element: $totalverb{$element}\n"; } } 27 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 28 appendix 3: sample data output from the perl script in appendix 2 child: adam search pattern: pro (don't) v vp-inf ........................................ age: 2;3.4 child want: 1 others have: 1 like: 5 want: 8 child data *chi: i [?] wan(t) (t)a [?] box . others data *mot: adam # want to close the box ? *mot: i guess she might like to see that . *mot: do you want to read a book ? *mot: do you want to see what i have ? *mot: do you want to write on here ? *mot: why do you like to throw your book ? *mot: would you like to have your books on the bookshelf too ? *mot: wouldn't you like to pick these up ? *mot: you like to walk ? *urs: do you want to bring over the high stool for me to sit on ? *urs: do you want to paste it ? *urs: do you want to play with them ? *urs: do you want to see this ? *urs: you'll have to pick them up . total verb count for child get: 48 like: 18 want: 3 ........................................ age: 4;10 child need: 1 28 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 a corpus study on the item-based nature of early grammar acquisition 29 want: 12 others have: 2 want: 5 child data *chi: i don't wan(t) (t)a do it . *chi: i don't wan(t) (t)a go under the bottom of the sea . *chi: i don't want to be a monkey . *chi: i don't want to be staying here . *chi: i wan(t) (t)a be my own self . *chi: i wan(t) (t)a go with you . *chi: i wan(t) (t)a make a kite . *chi: i wan(t) (t)a see it go higher . *chi: i want to go too . *chi: i want to try one out . *chi: ursula # i wan(t) (t)a make my kite . *chi: in case you want to go someplace ? *chi: we don't need to use de tops . others data *mot: paul # we might have to go upstairs # hmm ? *mot: you want to put that one on ? *urs: i think they have to be the double kind . *urs: d(o) you want to use another color ? *urs: d(o) you want to use the newspaper ? *urs: in case you want to carry it with you . *urs: you want to look at the little book . total verb count for child get: 43 have: 34 like: 23 need: 9 want: 34 ........................................ age: 4;10 child have: 8 like: 1 want: 8 others have: 5 like: 1 want: 4 29 hodges et al.: a corpus study on the item-based nature of early grammar acquisition published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 30 child data *chi: daddy # you don't have to be strong . *chi: i don't have to need it off . *chi: i wan(t) (t)a make a bow and arrow . *chi: i want to get some +... *chi: i want to hold it . *chi: after you finish # so nobody won't have to clean the place up . *chi: have to tear it out because how can we do it ? *chi: here # we going to have to build one with another string on it . *chi: hey # i wan(t) (t)a make a chair and a hammer . *chi: hey # you wan(t) (t)a make it a little smaller again ? *chi: if you want to try one # wait . *chi: ok # i'm gonna have to pick things up . *chi: see if the flowers would like to watch me . *chi: we gonna have to need a small bowl . *chi: you don't have to push him . *chi: you want to keep the hand+buckler@c ? *chi: you want to take only one home ? others data *mot: i don't want to have a dog . *mot: now i don't have to get any birthday presents . *urs: alright # but first i have to get it all collected . *urs: d(o) you want to ask your [?] mother [?] for that ? *urs: d(o) you want to look at them ? *urs: maybe you'll have to show me some things # alright ? *urs: then we have to make a sprinkling can . *urs: we may have to put another drop of ink on there . *urs: would you like to bring your chair over here ? *urs: you want to make some too ? total verb count for child get: 68 have: 51 like: 32 need: 9 want: 36 30 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/2 doi: https://doi.org/10.25810/ekag-3410 colorado research in linguistics 6-2004 a corpus study on the item-based nature of early grammar acquisition adam hodges valerie krugler deborah law recommended citation introduction initial study with pooled data from seven corpora verb follow-up longitudinal study of a single child general discussion summary customizing the lingo grammar matrix morphology customizing the lingo grammar matrix morphology sarah r. moeller university of colorado boulder the lingo grammar matrix is rapid development grammatical analyzer that is enabled by a customization system, including an online questionnaire. the questionnaire allows any linguist to quickly build a starter grammar for any language. however, its documentation is not easily navigable. this paper describes the process of customizing the morphological section of the grammar, using lezgi case inflection as an example. it outlines how suggestions for helping potential users to gain confidence in using the matrix. keywords: computational linguistics, grammatical analysis, lezgi, morphology, linggo grammar matrix 1. introduction the lingo grammar matrix (bender, flickinger & oepen 2002; bender et al. 2010) is “an open-source online repository of grammatical analyses which facilitates the rapid development of linguistically-motivated deep grammars compatible” (bender 2014: 2447). this rapid development is enabled by a customization system, including an online questionnaire. the questionnaire allows any linguist to quickly build a starter grammar for any language, but its documentation is not easily navigable. easy-to-read instructions could encourage linguists to begin implementing and testing the lingo grammar matrix on more languages. this paper describes the lingo grammar matrix and how to use its questionnaire (sections 2 and 3) with a case study on lezgi case morphology (section 4). section 5 outlines a presentation that would allow potential users to gain confidence in the customization system more quickly than is now possible. 2. the lingo grammar matrix the lingo grammar matrix (lgm) “combines a core grammar providing constraints and structures which are cross-linguistically useful with a series of libraries of analyses of crosslinguistically variable phenomena” (bender 2014: 2447). the core grammar consists of lexical rules believed to describe universal linguistic phenomena such as case marking or basic word order. the phenomenon-specific libraries cover language-specific grammatical rules. grammar 1 moeller: customizing lingo grammar matrix published by cu scholar, 2019 rules and semantics are conveyed in head-driven phrase structure grammar (hpsg) and minimal recursion semantics (mrs) and represented in a machine-readable language called type definition language (tdl) and displayed in the linguistic knowledge builder (lkb), a software environment that essentially serves as a graphic user interface. users can customize libraries by refining the core grammar rules. the customization process does not directly change the core grammar. instead, new rules are created which take a core grammar rule as “supertype”. the new rule inherits all features of the supertype and extends or constrains it. for example, customizing for strict sov word order would create a new rule that inherits all constrains the “supertype” word order rule to allow only sov word order. the customization system builds a deep precision grammar for any language via two steps. the first step builds a “starter grammar” using a customization system questionnaire (bender 2014; goodman 2013). completing the questionnaire produces a small implemented “starter grammar” that describes most phenomena encountered in simple main clauses such as basic word order, case marking, subject agreement, and tam marking. the second step begins once the pages of the questionnaire have been completed. to extend the grammar to phenomena not handled by the questionnaire, such as word order of questions, the tdl files must be edited directly. figure 1. lingo grammar matrix steps towards a precision grammar 3. the customization questionnaire lgm’s customization questionnaire is an online interface. its pages dedicated to topics such as word order, number, case, and “other features” guide users to provide a high-level description 2 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/2 doi: http://dx.doi.org/10.33011/cril.24.1.2 of a language’s main clause structure. for example, the first part of the case page presents a choice of nine case marking systems, including “no case”, nominative-accusative, and ergativeabsolutive, and the second part builds an expandable list of any other cases. the two parts together are added to the list of semantic/syntactic features that can be assigned to lexical items on the lexicon page or to morphemes on the morphology page. the morphology page is the most complex part of the questionnaire. the lgm treats morphology inferentially (goodman 2013: 5–6), defining relationships between roots and their inflected forms via morphological rules. the morphological rules are constructed in a three-level hierarchy: position classes, lexical rule types, and lexical rule instances. position classes group morphemes by the slot they occupy on the inflected word. within position classes, morphemes with common syntactic/semantic features or morphotactic restrictions are grouped in lexical rule types. a rule type may serve as supertype to other rules that will inherit the supertype’s features and morphotactics but may have additional features or constraints. these features and constraints do not target individual orthographic representations (essentially spelling changes such as the double consonants added to some english verbs after –ing: hop à hopping). orthographic idiosyncracies are specified by lexical rule instances which inherit all features from its lexical rule type. 4. customizing the lingo grammar matrix for lezgi lezgi (or lezgian) [lez] is a lezgic language belonging to the nakh-daghestanian family (simons & fennig 2018). it is spoken by over 600,000 people in the daghestan republic of russia and azerbaijan. like many languages in the caucasus mountains, lezgi exhibits ergative morphology as shown in sentence (1) and (2). (1) itim-di sev gata-na man-erg bear.abs beat-aorist ‘the man beat the bear.’ (2) sev itim-di gata-na bear.abs man-erg beat-aorist ‘the man beat the bear.’ 3 moeller: customizing lingo grammar matrix published by cu scholar, 2019 4.1 lezgi case lezgi has fourteen cases marked by suffixes. absolutive case is unmarked and the ergative suffix attaches directly to the noun stem. other cases incrementally add suffixes to the ergative (what haspelmath (1993: 74) calls the oblique stem). the genitive, dative, postessive, subessive, superessive, and inessive add one suffix (the inessive can be analyzed as a null suffix). the elative and directive meaning are added to the last four suffixes as a third suffix. table 1 illustrates this pattern. table 1. lezgi case inflection ‘bear’ ‘man’ absolutive sev itim ‘the n’ ergative sev-re itim-di ‘the n’ genitive sev-re-n itim-di-n ‘of the n’ dative sev-re-z itim-di-z ‘to the n’ addessive (adess) sev-re-v itim-di-v ‘at the n’ adelative (adel) sev-re-v-ay itim-di-v-ay ‘from the n’ addirective (addir) sev-re-v-di itim-di-v-di ‘toward the n’ postessive (poess) sev-re-q itim-di-q ‘behind the n’ postelative (poel) sev-re-q-ay itim-di-q-ay ‘from behind the n’ postdirective (podir) sev-re-q-di itim-di-q-di ‘to behind the n’ subessive (sbess) sev-re-k itim-di-k ‘under the n’ subelative (sbel) sev-re-k-ay itim-di-k-ay ‘from under the n’ subdirective (sbdir) sev-re-k-di itim-di-k-di ‘to under the n’ superessive (spess) sev-re-l itim-di-l ‘on the n’ superelative (spel) sev-re-l-ay itim-di-l-ay ‘off the n’ superdirective (spdir) sev-re-l-di itim-di-l-di ‘onto the n’ inessive (iness) sev-re itim-di ‘in the n’ inelative (inel) sev-rə-y itim-di-y ‘out of the n’ table 2 shows a position class chart for the nouns in table 1, including number suffixes that attach directly to the noun stem. 4 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/2 doi: http://dx.doi.org/10.33011/cril.24.1.2 table 2. position class chart for -re and -di nouns 0 stem +1 number +2 case1 (+3) case2 (+4) case3 -ø ‘sg’ -ar ‘pl’ -ø ‘abs’ -re/di ‘erg’ -n ‘gen’ -z ‘dat’ -v ‘adess’ -q ‘poess’ -k ‘sbess’ -l ‘sress’ -ø ‘iness’ -ay ‘ad/po/sb/sr/in -el’ -di ‘ad/po/sb/sr -dir’ table 2 simplifies lezgi noun morphology in two ways. first, morphophonological changes that produce two plural suffixes and the inelative are ignored because the lgm does not handle morphophonology (goodman 2013: 11). users are expected to provide underlying phonological representations. second, the ergative case has eight more suffixes. although lezgi is described as having no noun classes and no agreement (haspelmath 1993; manning 1994), these suffixes are selected by semantic and phonological factors that indicate possible remnants of noun class agreement. for example, the -a ergative suffix occurs on personal names ending in a consonant, but the –i ergative suffix occurs on the same word if it is used as a personal name (e.g. cükwer-a ‘of the flowers’ but cükwer-i ‘flower’s’). 4.2 customizing morphological rules in order to describe the lezgi case morphology in table 2 three pages of the questionnaire were modified. first, an ergative-absolutive system was chosen and the other twelve cases were declared on the case page. second, noun roots were added to the lexicon as two noun types: one that takes -re ‘ergative’ suffix and one that takes -di ‘ergative’. it is worth noting that since the lexicon allows semantic/syntactic features to be assigned to noun types a new user might mistakenly assign the absolutive case as a feature of bare noun roots. subsequently adding a ergative suffix would give the noun two clashing case features, which would prevent parsing because of lgm incremental approach to morphology (goodman 2013: 5–6). instead, the absolutive case should be added by morphological rule as a zero affix. finally, on the morphology page, the case suffixes were defined and morphological rules created to describe how they attach to noun roots. a position class (pc) was created for number and for ergative and absolutive case. as figure 2 shows, the case (pc2) takes number (pc1) as its 5 moeller: customizing lingo grammar matrix published by cu scholar, 2019 input and is obligatorily attached as a suffix. since the number suffix attaches obligatorily to all nouns, together the position classes dictate that a noun must be inflected for both number and case, in that order. figure 2. lezgi case position class the singular and plural morphemes are defined in the pc1 as two lexical rule types each with one lexical rule instance. similarly, in pc2, the absolutive case is defined in a lexical rule type with one lexical rule instance (“no affix”). the ergative case is defined as a hierarchy of three lexical rule types. the first rule (noun-pc2_lrt1) serves as the supertype and its semantic feature (ergative case) is inherited by the other two: erg-di (noun-pc2_lrt3) and erg-re (nounpc2_lrt4). these two rules each provide a lexical rule instance: di and re. the two subrules differ only in morphotactic constraints. the constraints on erg-di (noun-pc2_lrt3) enforces cooccurrence with the -di noun type in the lexicon, as figure 3 shows, and erg-re (nounpc2_lrt4) enforces co-occurrence with -re nouns. 6 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/2 doi: http://dx.doi.org/10.33011/cril.24.1.2 figure 3. ergative -di lexical rule type morphotactic constraints were intended to prevent underinflection and overinflection (goodman 2013: 7, 32) and this can be tested in lkb by generation. putting the morpheme variants in a hierarchy of lexical rule types eliminates overinflection, as shown in figure 4 (cf. figure 5). figure 4. correct case morphology generation if a new user did not understand the “supertype” hierarchy and instead represented ergative case as two lexical rule instances in a single lexical rule type, the generation would yield the overinflection in figure 5. the suffix -di should only appear on itim ‘man’ (and -re only on sev ‘bear’). figure 5. incorrect case morphology generation 7 moeller: customizing lingo grammar matrix published by cu scholar, 2019 except for bipartite verb stems, morphotactic constraints cannot target lexical rule instances via the questionnaire. this limitation raises two questions. first, why allow multiple lexical rule instances in one lexical rule type in the questionnaire? except in rare cases of free variation, it seems unlikely that orthographic variants can be properly handled by one lexical rule type. at the same time, using multiple lexical rule instances in a single lexical rule type would seem to new users a deceptively simple solution for handling morphophonology or agreement paradigms. second, what advantage does the supertype hierarchy give over unrelated lexical rule types? creating two lexical rule types for the ergative case and assigning them the same syntactic/semantic feature produces the same results. morphotactic constraints would still forbid ergative suffixes to co-occur on the same noun and overinflection with the absolutive case is disallowed because all the case rule types belong to the same position class. neither the questionnaire nor goodman (2013) provide insights into why one configuration might be preferred. 4.3 lezgi case morphology: further steps since the lgm takes an incremental-inferential approach to morphology, attempting to define all columns in table 2 as position classes would cause each additional suffix to clash with the case feature assigned to the previous suffix. handling the remaining 12 cases in lezgi requires a decision whether to reflect the morphology in table 2 or simplify the morphological rules. instead of defining all 14 cases on the case page, the oblique cases could be defined on the other features page. this would create new semantic “case” features that would not clash with the erg-abs position class. this approach would reflect lezgi’s sequence of case suffixes. a second approach would treat each sequence of suffixes as a single morpheme. each case “morpheme” would be added as a lexical rule type to the second position class. this approach does not accurately describe lezgi morphology, but it is simpler. as goodman (2013: 3) points out, the goal of computational grammars is to “parse and generate valid sentences” rather than “capture the behavior of interesting linguistic phenomena.” 5. towards greater accessibility lgm’s developers wish to collect language specific libraries (bender 2014) so moving the state of the art forward means more data from more languages. customizing a precision grammar 8 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/2 doi: http://dx.doi.org/10.33011/cril.24.1.2 requires a great deal of time and some specialized knowledge. by eliminating the need to know hpsg and tdl, lgm’s customization questionnaire has reduced the learning curve from perhaps 80+ workhours to about 30. however, much of the learning time must still be spent finding and identifying resources for learning the customization process. on the lgm homepage, publications are sorted chronologically, not by relevance to learners. the wiki provides some instructions but not in logical sequence. certain instructions lack screenshots that would allow new users to compare their progress against desired results. other instructions appear to refer to previous versions of the lkb interface. online courses that teach the lgm assume a background in grammar engineering or hpsg. this section assembles resources from literature about the lgm, online course materials, and the experience of the author as a new user. it summarizes basic information that is needed when beginning to customize the lgm. it is imagined as an “orientation” page for the questionnaire, allowing potential users to assess what resources they need gather and what steps they will take as they proceed. “before you start” these instructions assume you are a trained linguist. it does not assume you are familiar with computational linguistics, hpsg, or mrs. it does assume you are somewhat familiar with unix. overview to customizing a starter grammar there are two presteps to building a starter grammar with the lgm’s customization questionnaire: 1) download the software, 2), choose a language and fill out the general information page. the rest of the process is an iterative cycle that will continue until all pages in the questionnaire have been completed and tested. overview of customization cycle 1. describe one to three phenomena by filling out the relevant pages of questionnaire. 2. add grammatical and ungrammatical sentences that illustrate the phenomena to your test suite. 3. on the lexicon page, enter any new vocabulary in those sentences. 4. on the morphology page, add/edit position classes, lexical rule types (morphemes sharing semantic/syntactic features and morphotactic constraints), and lexical rule 9 moeller: customizing lingo grammar matrix published by cu scholar, 2019 instances (orthographic representation of morphemes) to cover all new morphology in the test suite. 5. download and save latest version of the choices file. 6. generate a new version of the starter grammar. 7. unzip the grammar and load it into lkb software. 8. try parsing some individual sentences from your test suite. 9. generate sentences to examine morphology or try batch parsing the test suite. 10. based on generation and batch parsing results, edit the questionnaire. 11. repeat steps #6-11 until satisfied, then proceed to step #1 in order to add new phenomena. necessary resources • reference grammar for your chosen language • software • some knowledge about hpsg and mrs • understanding of the lgm’s approach to morphology downloading software • list of software: o lkb – this is where you test the latest version of your grammar; it parses and generates sentences, displays phrase structure trees, attribute value matrices (avm), and semantic representations. it only runs in ubuntu. o virtual box – allows you to run the ubuntu operating system inside your windows or os x computer. o ubuntu – you will download a package called ubuntu+lkb that has lkb already installed in it. o emacs – this program already installed in ubuntu. it launches lkb and sometimes displays useful information while lkb is running (e.g. debug report). it has a built-in tutorial that you may find useful. however, you really only need one command: m-x lkb (type alt+x, then lkb) to launch lkb. • downloading software o it is not uncommon to run into complications while downloading software. this can be a discouraging way to start, so have someone nearby who can help. o go to: http://depts.washington.edu/uwcl/twiki/bin/view.cgi/main/knoppixlkbvboxapp. follow the instructions under “install virtualbox” and “set up the ubuntu+lkb appliance”. o create or choose a folder to keep related files and follow the instructions for “setting up a shared folder” at the bottom of the link. 10 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/2 doi: http://dx.doi.org/10.33011/cril.24.1.2 • becoming familiar with lkb o take time to play around with lkb and become comfortable with it. go to: http://courses.washington.edu/ling567/lab1.html. follow the instructions under “grammar customization: get a small grammar for english” and “lkb: getting started” (ignore the first step). if you are not familiar with ubuntu, we highly recommend that you find someone to guide you through these steps the first time. you will repeat them often. learning hpsg and mrs it is possible to customize a starter grammar without knowing hpsg or mrs. it is difficult, however, to identify syntactic or semantic problems and impossible to go beyond the questionnaire if you cannot decipher the avm used to display hpsg and mrs in lkb. • hpsg introduction: levine, robert d. 2003. head-driven phrase structure grammar. encyclopedia of cognitive science. • avm cheat sheet: o synsem – contains loc and nonloc properties o loc(al) – locally relevant properties of the word, phrase, or clause; identifies lexical or phrasal properties via cat o nonloc(al) – information that extends over a larger syntactic domain, (e.g. long distance dependency) o cat – contains val and head values o val(ence) – valence information; specifies subject as subj, non-subject arguments as comps o head – properties of lexical items shared by the phrase it heads, (e.g. (pos), case, aux(iliary), etc.) o arg-st(ructure) – verb's argument structure; identical to the list of comps + subject o spr – modifier/specifier o conx – contextual information o c-cont or cont(ent) – logical aspects of semantic interpretation; relations among agent, patient, and so on o index – part of the cont for nominals; encodes reference mrs reference guide (somewhat out of date): flickinger, dan, emily m. bender and stephan oepen. 2003. mrs in the lingo grammar matrix: a practical user's guide. ms. • general introduction: copestake, ann, dan flickinger, carl pollard & ivan a. sag. 2005. minimal recursion semantics: an introduction. research on language and computation 3.281–332. 11 moeller: customizing lingo grammar matrix published by cu scholar, 2019 representing morphology the morphology page will describe your chosen language’s morphological rules. this is the most complex part of the questionnaire, but once you learn the basic principles, it is fairly simple to use. • lgm morphology overview with some examples (read all of sections 2 & 3): goodman, michael wayne. 2013. generation of machine-readable morphological rules from human readable input. (ed.) sanghoun song & joshua crowgey. uw working papers in linguistics 30. http://depts.washington.edu/uwwpl/vol30/goodman_2013.pdf (23 september, 2016) getting started start by taking notes on the language’s grammar and constructing a small test suite of example sentences. follow these instructions: http://hpsg.stanford.edu/05inst/prep.html. notes on test suite as you proceed, you will need to test your starter grammar against real language data, so construct sentences that demonstrate and violate the phenomena described in the questionnaire. it is wise to start with simple sentences and reuse words as much as possible. avoid nonverbal predicates and copulae until later (e.g., the cat is old.). here are some sample sentences: http://moin.delph-in.net/matrixmrstestsuite. here, with examples, are all grammatical phenomena that the test suite should eventually cover: http://compling.hss.ntu.edu.sg/courses/hg7021/testsuites.html#phenomena. skim this page and start thinking about how to illustrate these phenomena using minimal vocabulary and morphology. • test suite specifications o if the language is not normally written with spaces between words, add spaces. o all examples should be complete sentences. o each sentence should be on a new line. o ungrammatical examples should generally have only one thing wrong with them. o since the system is not designed to handle morphophonology, choose inflections that do not undergo morphophonological changes or else only write underlying phonemic representations. o save your test suite as a plain text file (.txt). o encode the file as utf-8. 12 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/2 doi: http://dx.doi.org/10.33011/cril.24.1.2 notes on test suite format if you wish your starter grammar to be available online for future users, the test suite needs to be formatted as described here: http://compling.hss.ntu.edu.sg/courses/hg7021/testsuites.html#formatting. otherwise, simply type the example sentences directly into the text file and put a semicolon (;) or number sign (#) before any comments (for yourself or others). for more information for more help, explore the links on the lgm wiki page: http://moin.delph-in.net/matrixtop. 6. conclusion the lgm encourages the implementation of computational grammars. its customization questionnaire provides a relatively simple way to build a precision grammar for any language, including lesser resourced languages, as this case study with lezgi demonstrates. however, getting started proves difficult because the questionnaire’s supporting resources are not organized for independent learners. this might be solved by more prominent and better organized instructions for new users such as those outlined in this paper. references bender, emily m. 2014. language collage: grammatical description with the lingo grammar matrix. proceedings of the ninth international conference of language resources and evaluation (lrec-2014), 2447–2451. http://www.lrecconf.org/proceedings/lrec2014/pdf/639_paper.pdf. bender, emily m., scott drellishak, antske fokkens, laurie poulson & safiyyah saleem. 2010. grammar customization. research on language and computation 8(1). 23–72. doi:10.1007/s11168-010-9070-1. bender, emily m., dan flickinger & stephan oepen. 2002. the grammar matrix: an opensource starter-kit for the rapid development of cross-linguistically consistent broad-coverage precision grammars. proceedings of the 2002 workshop on grammar engineering and evaluation-volume 15, 1–7. association for computational linguistics. http://dl.acm.org/citation.cfm?id=1118785 (19 october, 2016). goodman, michael wayne. 2013. generation of machine-readable morphological rules from human-readable input. (ed.) sanghoun song & joshua crowgey. uw working papers in 13 moeller: customizing lingo grammar matrix published by cu scholar, 2019 linguistics 30. http://depts.washington.edu/uwwpl/vol30/goodman_2013.pdf (23 september, 2016). haspelmath, martin. 1993. a grammar of lezgian. berlin; new york: mouton de gruyter. manning, christopher. 1994. ergativity: argument structure and grammatical relations. stanford phd. http://www.press.uchicago.edu/ucp/books/book/distributed/e/bo3643756.html (30 june, 2014). simons, gary f. & charles d. fennig (eds.). 2018. ethnologue: languages of the world. twenty-first. dallas, texas: sil international. http://www.ethnologue.com. 14 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/2 doi: http://dx.doi.org/10.33011/cril.24.1.2 colorado research in linguistics 6-2019 customizing the lingo grammar matrix morphology sarah r. moeller recommended citation microsoft word moeller-cril2019-final.docx loss and consequence: an examination of the old english case marking system as opposed to that of other old germanic languages loss and consequence: an examination of the old english case marking system as opposed to that of other old germanic languages* * denise e. walters university of colorado at boulder old english—early in its existence—did not differ much from its germanic cousins. in fact, these languages could be considered distant dialects from one another. however, through the course of its development, old english lost part of its germanic morphology: the case-marking system. the loss of this system had such an impact on the development of the language that the results are seen in modern english. this paper examines these results, and to an extent the reasons, behind this reduction. introduction “linguistic change is initiated by speakers, not by languages” (milroy 1997: 311). most linguists understand the truth behind this statement. one must take into consideration the speakers’ usage when considering language changes, regardless of the examination being undertaken. in the case of a diachronic study—even, or maybe especially, a typological one—the scholar must not only deal with the data, but also the relationship between that language and its cousins. otherwise, a complete understanding might not be reached. for example, for the purposes of this study, old english will be analyzed with relation to its germanic cousins or sisters. extensive changes in morphology occurred in late old english (oe) and in early middle english (eme). yet for various reasons, these changes did not occur within english’s germanic cousins. the most obvious change was that the old system of case marking was nearly completely swept away. this drastic reduction of case marking has usually been seen as responsible for many syntactic changes, as well. these reductions are quite visible, as are the consequences of these reductions. before the reductions occurred, english—in its infancy—looked so similar to its germanic cousins that while the languages could not be considered dialects or variations1 of the same language, they were definitely not mutually unintelligible. 1. the background before examining the grammar of the nouns in oe that later underwent a change in case assignment and/or grammatical relations, it is necessary to make some general observations about oe syntax. in oe, case marking played an important role in signaling * i would like to thank regina pustet for her encouragement during her typology course. 1. for length reasons, these issues will not be addressed in this paper. colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by denise walters. 1 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 grammatical relations.2 four cases were productive in oe: nominative, accusative, dative, and genitive.3 the latter three cases could all be used without a preposition to mark objects of verbs. the accusative was always used with verbs inherently high in transitivity (andrews 1985); however, case marking is harder to predict with verbs that hopper and thompson (1980) consider less transitive. moreover, many verbs show variability in their case marking. most verbs which could take genitive objects also sometimes appear with objects in another case; the alternation between genitive and dative, as with gehelpan ‘to help’, is less common than that of accusative and genitive, as with afandian ‘to test, prove’ and abidan ‘to wait for, await’ (bean 1983: 37). afandian usually took a genitive object when it meant ‘test’ and an accusative object when it meant ‘prove.’ however, the accusative case was sometimes extended to the ‘test’ meaning in examples where it seems very difficult to argue for a difference in meaning which would explain the accusative (bean 1983: 39). it is likewise with abidan: ‘to wait for’ requires the dative and ‘to await’, although semantically similar, requires the accusative (bean 1983: 39). it is generally assumed amongst scholars that at least some case markings must be lexically specified, although these lexical specifications may well follow certain patterns based on semantics. most current syntactic theories assume the existence of two distinct types of case marking: lexical (or inherent) and structural (or syntactic) case marking. structural case marking is the default while lexical case is assigned idiosyncratically in lexical entries.4 lexical case differs from syntactic case in that syntactic processes do not affect it. for example consider the difference between (1) a and (1) b in icelandic: (1) a. strákarnir voru kitlaðir boys-the-nom-pl were tickled-nom-masc-pl ‘the boys were tickled.’ b. stráunum var bjargað boy-the-dat-pl was-sg rescued-nom-masc-sg ‘the boys were rescued.’ (helfenstein 1870: 281) these sentences differ both in the surface case of the boys and in the agreement or lack of it of the passive participle and the auxiliary verb. the most widely accepted explanation for this difference is that the boys receives case structurally with kitla, but gets its case marking lexically from bjarga, which requires a dative object (helfenstein 2. it must be noted, however, that even at the oe stage, a good deal of syncretism had crept into the casemarking system, due mainly to phonological change. for example, by the oe stage, the distinction between the nominative and accusative forms had been lost in a large class of nouns, the masculine a-stem, in both the singular and the plural (allen 1995: 25). 3. additionally, there was an instrumental case, but within prose, it had been almost completely replaced by prepositions, except within certain expressions. the instrumental was mainly retained in poetry and religious texts. 4. precisely how structural case marking works will depend on the theory adopted. for example, koopman and sportiche (1990) suggest that there is more than one type of structural case marking; however, that will not be discussed here. it was merely mentioned to explain that different theories support different claims. 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 3 1870: 282). the lexical entry of kitla does not impose any case marking on the arguments associated with the verb, and so these arguments get their case marking structurally based on their syntactic role at the surface level. oe was similar to modern icelandic in having a difference between lexical and structural case, although the case-marking patterns of the languages differ in some interesting respects, mainly due to the different principles followed in the application of case markings. however, at this point, it should be noted that dative and genitive case were regularly preserved under passivization in oe, as they are in icelandic. for example, deman ‘to judge’ takes a dative object and the dative case remains when the verb is passivized: (2) hi ne demað nanum men, ac him they not judge no-pl-dat men-pl-dat but them-dat bið gedemed is-sg judged. ‘they will not judge any men, but they will be judged.’ (ælc.p.xi.369) however, the derived subjects of verbs which take accusative objects show up in the nominative case. the simplest explanation is that deman undergoes passivization just like verbs taking accusative objects, but because the underlying object of the verb receives dative case lexically, it does not receive nominative case structurally when becomes the subject. being a case-marking language, oe signaled its grammatical relations by inflection, rather than constituent order, making possible much greater variation in the order of constituents than is found in the modern language. however, oe constituent order was not by any means completely free, and case marking was far from unambiguous. it is widely agreed that discourse factors played a very important role in oe constituent order; for example, most scholars concur that a noun phrase which had been mentioned before, and thus was ‘old information’ was likely to be placed near the front of a sentence.5 there is less agreement on the question of what role, if any, syntactic categories such as subject, verb, and object played. the most common argument is that both grammatical categories and discourse factors played an important role in oe constituent order, at least by the late oe period, as represented by ælfric. there is one final aspect of oe syntax that must be talked about for background purposes: the use of the formal subject hit ‘it’. in contrast to modern english (mode), it is possible in oe for tensed clauses to appear with no np in the nominative case. it is not difficult to find examples in oe in which a verb appears with a sentential complement and there is no anticipatory formal subject, which would be obligatory in mode: (3) đa gelamp þæt he… then happened that he ‘then it happened that he…’ (bede qtd. in healey and venezky 1980: 232). 5. helfesntein 1870; healey and venezky 1980; gneuss 1996 3 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 it has been argued6 that formal hit was not used in the above sentence because its purpose in oe was simply to preserve verb-second order, and it was, therefore, not necessary when an adverb appeared before the verb. however, allen (1986a) determined that this formal subject appears frequently in sentence in which it is not necessary for maintaining verb-second order. it is quite easy to find examples like (4), in which the verb would be in second position even with the addition of hit, and like (5), where its presence ensures that the verb is in third,7 rather than second, position: (4) þa gelamp hit þæt æt ðam gyftum… then happened it that at the wedding ‘then it happened that at the wedding…’ (ælc.th.p.569) (5) on ðære tide iu hit getimode swa,… þæt he stod… in the time before it happened so,… that he stood… ‘at the earlier time it happened so…that he stood…’ (ælc.p.xiv.i) a concordance corpus count of the verb gelimpan ‘to happen’ used with a sentential complement shows that the formal subject hit in fact appears in 83 percent of the examples in which it is ‘not necessary’ to maintain verb-second order, as in (4) and (5). of the examples in which the formal subject is used, 71 percent do not have verb-second order (allen 1986a: 468). it does appear, however, that hit was used to prevent verb-first order, since the placeholder is nearly always used when the verb would otherwise have been initial. thus, formal subjects were greatly preferred in this sort of sentence in oe, regardless of whether the verb was in second or further position. the only positional constraint which played a role in the use of hit was the general—but not total—prohibition against verb-first declarative sentences. 2. the case markings of old english: explanation and reduction8 having established the background that informed the morphological changes, and therefore the syntactic changes, of oe, it is now time to talk about the changes themselves. the loss of case-marking distinctions in english has generally been seen as responsible for profound changes in the language. compared with the present-day language, oe is highly inflectional. nouns have four cases and three genders [cf. appendix a for a brief summary of oe declensions]; verbs inflect for person and number and for the indicative and subjunctive moods. further, in 6. haiman 1974 7. that is, assuming that hit is to be treated as a separate constituent here, rather than as a clitic. if it treated as a clitic, then the verb is still in second position. 8. for purposes of this paper, only the case marking system of nouns and adjectives are to be considered. 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 5 the oe noun phrase, there is agreement between noun and modifying adjective, much like the present-day german. the oe inflectional system derives directly from that in germanic. however, oe begins to show the loss and simplification of inflections which characterizes the later stages of english and which eventually creates a language with remarkably few inflections compared to its cousins. one change which is frequently regarded as important in the loss of the ‘impersonal’ constructions is the syncretism between nominative and dative nominal cases. in oe, the distinctive dative case marking on the nominal clearly denotes which word is the ‘impersonal’ verb (6)9, but once the dative case marking disappears from the nominal paradigm, it becomes impossible for language-learners to be certain if the nominal present should be analyzed as an object (albeit, an indirect one) or as a subject. (6) þam cyninge ofhreoweþ the king-dat to feel pity for something ‘the king pities…’ (walters) because the experiencer was in the preverbal position, which is typically occupied by the subject, the former object has been reanalyzed as a subject, despite the existence of examples with pronouns (7), in which the case marking is unambiguously nonnominative: (7) him ofhreoweþ he-dat to feel pity for something ‘he pities…’ (walters)10 the loss of the dative inflection for nouns was an instance of syncretism of forms, rather than loss of a category distinction, since the distinction between nominative and object case is still found in the pronouns. however, another change which has taken place, the collapse of the distinction between accusative and dative cases, involves the loss of an important category distinction, (8) and (9). the loss of this category distinction profoundly affects the case marking system of oe: the distinction between a lexically assigned case and a structurally assigned case is wiped out. (8) and him gelicade hire þeawas and þancode gode and him like her virtues-nom/acc and thanked god ‘her virtues pleased him, and he thanked god…’ 9. the nominal under discussion is underlined, and the verb is bolded. 10. the dative/nominative syncretism which occurred in the nominal system has also been regarded as the trigger for another sweeping syntactic change: the introduction of a new passive. for length reasons, this will not be discussed in this paper [cf. allen, c. 1995. case marking and reanalysis: grammatical relations from old to early modern english. oxford: clarendon press. ] the loss of this distinction is supposed to have led to unclarity about the grammatical relations involved in passives such as the king was given a gift and the king was harmed. in both of these passives, the king would have been clearly dative in oe, but was liable to reanalysis as the nominative subject at a later stage. 5 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 or ‘he liked her virtues, and thanked god…’ (healey and venezky 1980: 352) (9) þonne soðilce gode licað ure drohtnunge, þonne we þa then truly god-dat likes our living, when we the god, þe we onginnað, on urhwuniendum end gefyllað good, which we begin, in preserving end fulfill ‘then truly does our way of life please god, when we carry through to the end the good which we have begun’ or ‘then truly does god like our way of life when we…’ (healey and venezky 1980: 120). as in most germanic languages—or even languages from other branches of indoeuropean—oe has four major types of vocalic nouns (nouns with vowels at the end of the stems): the a-stems, the ō-stems, the i-stems, and the u-stems, of which the first two are by far the most common. the a-stems are frequently referred to as the ‘masculine astems’, whose typical paradigm is represented by table 1. table 1. declension of the masculine a-stems in oe. example: stān ‘stone’ singular plural nominative (nom) stān stānas accusative (acc) stān stānas genitive (gen) stānes stāna dative (dat) stāne stānum the a-stems provide some good examples of how even at the earliest recorded stage of english, considerable syncretism of form—compared to the forms which can be reconstructed for proto-germanic—have taken place in english. for example, the nominative and accusative singular forms are distinct in the proto-germanic language— and even in some of the other germanic languages—but have fallen together by the earliest oe stage. the germanic nominative singular form ended in –az, and the accusative form ended in –am. both these forms disappeared by purely phonological processes which affected unaccented syllables in pre-oe [cf. campbell 1959: 570], so that the nominative and accusative form for ‘stone’ converged on stān, with no suffix [cf. icelandic harmr ‘harm’, in which z>r, but oe harm]. syncretism had taken place in the plural even before the oe stage, with the original –ôs ending of the nominative extended to the accusative by the west germanic stage (c.1100). syncretism is also found in a group of nouns known as the ‘feminine ō-stems’, but the categories affected are different. in oe, the accusative, genitive, and dative forms of these nouns all ended in –e. this suffix is the reflex of three separate suffixes at the germanic stage (campbell 1959: 586), as shown in table 2, which for length reasons only shows the singular paradigm. 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 7 table 2. declension of the feminine ō-stems in oe. example: giefu ‘gift’ (singular only) proto-germanic form of suffix nom gief-u -ō acc gief-e -ōm gen gief-e -ôz dat gief-e -ai note: if stem is long, the –u of the nominative singular is deleted, as in lār ‘lore, doctrine’. campbell indicates that the development to –e is regular by phonological processes, except in the genitive, where paradigmatic pressure and analogy seem to have played a role (1959: 586). non-phonological pressures also play a role in the syncretism which took place in the nominative and accusative plurals of the feminine ō-stems. in germanic, these suffixes were –ôz (nom) and –ōns (acc) (lightfoot 2002: 99). the reflexes of these are –a and –e, respectively. one of the phonological processes contributing to much of the syncretism that took place by the end of the oe period was the reduction in the variety of vowels found in final unstressed syllables. the distinction between the back vowels in this environment was already showing clear signs of weakening in the kentish charters of the 9th century, and was completed in northumbrian in the 10th century, according to campbell (1959: 377). the vowels seem to have remained distinct in west-saxon for a longer period, although confusion of /a/ and /o/ is also apparent from scribal errors in early west-saxon (lightfoot 2002: 95). the front vowels were also affected, with /æ/, /e/, and /i/ falling together in a symbol written as at an early date (campbell 1959: 369). this means that by the late oe stage, the only distinction in the vowels of suffixes was between higher and lower vowels. this distinction disappeared in the 11th century, when the front and back vowels had ‘largely coalesced’ (campbell 1959: 379). this reduction of unstressed vowels meant that the case inflections were now less effective than they had been at reflecting category distinctions. none of the above mentioned changes resulted in the loss of a category distinction. for example, although with many nouns the nominative and accusative are identical in form, the category distinction between nominative and accusative is still prominent, being reflected in the forms of adjectives and demonstratives, as well as in the forms of the feminine nouns [appendix b. tables 7 and 8, respectively]. towards the end of the oe period, two phonological changes combine to have a devastating effect on the inflection of adjectives and determiners, as well as affecting the nominal declensions. the first change is a replacement of /m/ by /n/. the timing of this change with respect to other changes characteristic of eme has been documented by moore (1928: 242). on the basis of an examination of a large number of texts, moore concludes that this change was certainly completed by the end of the 11th century (1928: 261). according to campbell (1959: 378), the change is already evident in early westsaxon in the dative plural ending of nouns, which is occasionally found as –un rather than the expected –um by the beginning of the 11th century. 7 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 the above change is by itself enough to cause a considerable amount of syncretism, especially since it was combined with the reduction of the vowels. for example, the dative plural –um was no longer distinct from the –an ending which was so frequent in the weak forms of the adjectives and also in a class of nouns, called ‘weak nouns’, which declined similarly to these adjectives. the syncretism becomes massive when another change quickly follows: the loss of final /n/ in unstressed syllables.11 this change affects the newly created /n/ as well as the old ones. for example, there is no longer any distinction between the nominative and accusative singular of any masculine nouns. this distinction has already disappeared in the strong a-stems, but in oe the distinction is maintained in the weak nouns (10). (10) hunta ‘hunter’ (nom) versus huntant (acc) (walters) earlier, the feminine weak nouns have had a distinction between –e in the nominative and –an in the accusative, but the endings combine when the nasal was lost and the vowels became identical. this means that some strong feminine nouns are the only ones which maintained the distinction between nominative and accusative singular (11). (11) dæ”d ‘deed’ (nom) versus dæ”de (acc) (walters) it seems likely that it was impossible to maintain the distinction between nominative and accusative feminine nouns once so few nouns show this distinction (moore 1928: 262). by the late 11th century, the –e of all the non-nominative singular forms of these feminine nouns is extended to the nominative. with the reduction of final vowels and the loss of the final nasal, these nouns become in effect indeclinable, and the formal distinction between nominative and accusative disappears for all nouns. at the end of the 11th century, the case-marking system of english is still intact as a system, in that the same case categories are involved, but the evidence supporting the category distinctions are now greatly reduced because of widespread syncretism of forms. in peterborough, located at the southern border of the northeast midlands area, it appears that the case-marking system at the end of the first third of the 12th century is not radically different from oe in the category distinctions which it makes, but syncretism of forms have brought the system very close to extinction. although the distinction between dative and accusative still maintains a tenuous hold, the distinction is no longer marked in the feminine or plural pronouns, or in the nouns, where the old dative inflection has now become reanalyzed as a post-prepositional inflection. verbal selection of genitive objects either have disappeared entirely or are at least unusual in this area. the final continuation12 shows that by the middle of the century, the dative/accusative distinction 11. final nasals were already lost in some morphological contexts, such as in the infinitive, in the early northumbrian texts. 12. the second lengthy addition added to the anglo-saxon chronicle, dealing with years 1132-1154. the first lengthy addition, dealing with years 1122-1131, is called the first continuation. 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 9 have entirely disappeared, and the distinction between subjects and objects is no longer marked in the determiner system, (12) through (14). (12) & benam ælc ðone riht hand and deprived each-un the-acc/dat right hand-un ‘and deprived each of them of their right hands’ (pc 1125.9) (13) him me hit beræfode him man it-acc/dat bereaved ‘he was deprived of it’ (pc 1124.51) (14) & iærnde ða þurh him & ðurh ealle his freond and asked then through him and through all his friends namcuþlice þone abbotrice namely the-acc/databbacy-acc/dat ‘and then asked, through him and all his friends, specifically for the abbacy.’ (pc 1127.49) in all these examples, the verb formerly required or at least allowed an object in the genitive case; for example, the verb of (13), benimen ‘deprive’, would have earlier normally had a deprive in the accusative or the dative, and an object of deprivation in the genitive, although a minority pattern with a dative deprive and an accusative object of deprivation already exists at the oe stage. these objects are now indistinguishable from ordinary direct objects, appearing either unmarked or in the old accusative form. not surprisingly, no genitive objects are found in the final continuation. further south, the dative/accusative distinction was still quite healthy even towards the end of the 12th century, although the category distinction is only optionally marked. objects are still frequently marked with genitive case. however, there is limited evidence available from the original texts from the southern part of the country to reach firm conclusions about how the loss of case marking proceeds. in view of the fact that only a few remnants of the old case-marking system are to be found in the final continuation of the peterborough chronicle around the middle of the 12th century, it is no surprise to find that most of these remnants have disappeared entirely by the beginning of the 13th century in the northeastern part of the county. the nominal suffix –e is still found fairly frequently in the area’s texts, namely the ormulum—a long poem written by a monk with the intention of explaining the gospels in english. a distinct dative form is hardly ever used for plural nouns, where a single form has usually been generalized to all cases, and such examples found are restricted to the objects of prepositions. with singular nouns, the most frequent use of the old dative suffix is again on the objects of prepositions, but the inflection is also found in some other positions, namely the complements of nouns and adjectives, (15) and (16): 9 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 (15) unþeaw gode laþest vice god-dat loathsome-est ‘a vice most loathsome to god’ (ormulum a f.54) (16) cristeto wurðmund christ-dat to glory ‘for the glory of christ’ (ormulum m f.25.10) however, such examples are not common, and most non-prepositional uses of the dative are to be analyzed as fixed expressions which can be listed in the lexicon. most importantly, there are no examples of the –e inflection on what is the equivalent of the indirect object in mode which are not preceded by a preposition. numerous examples show that the unmarked form could be used for the ‘indirect’ object as well as the direct, and that the indirect object never had dative inflection (17) and (18): (17) ha chepeð hire sawle þe chapmon of helle she sells her soul the merchant of hell ‘she sells her soul to the merchant of hell’ (aw 213.28) (18) ne talde ha þen engel na tale not told she the angel no tale ‘she did not tell the angel any tale’ (aw 35.30) despite the lack of case marking, the recipient and theme are not yet distinguished by a fixed word order. as has been evidenced, a distinct inflectional category which could be called ‘dative’ still exists, but this case is now an exclusively syntactic case which is not lexically selected by any verb. in the pronominal system, the distinction between the accusative and dative has been completely lost in the feminine, neuter,13 and plural pronouns. the dative forms hire (feminine) and ham (plural) have entirely supplanted the old accusative forms. in the neuter, however, it is the old accusative form which has replaced the dative form. although examples with a neuter dative are rare, the few examples which are to be found shown that the dative form has already been replaced by the accusative form (19): (19) nis hit neod zeorde? not-is it-un need rod? ‘it is (i.e. the child) not in need of a rod?’ (aw 167.23) 13. gender is by this time normally natural gender, not grammatical, although in this dialect few remnants of the old system in the use of feminine pronouns to refer to some nouns which historically have feminine gender are found. 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 11 this pronoun would have been in the dative case in this construction in oe. a distinction between the old accusative form of the neuter pronoun, hit, and the old dative form, him, has lasted longer than did the dative/accusative distinction with any other pronoun. if this distinction in forms continues to reflect a category distinction between accusative and dative case, it would be evidence that this distinction persists in the grammar much longer than suggested by previous scholarship [cf. allen 1995]. 3. the comparisons to english’s germanic cousins the loss of case-marking distinctions in oe is a surprisingly orderly and systematic affair, as it has been attempted to show. the loss of the accusative/dative distinction does not immediately result in the loss of all lexical case markings. proposed dative experiencers continue to flourish for a long period with the ‘impersonal’ verbs. however, unlike english, many of its germanic cousins did not suffer from this type of reduction, or at least they did not suffer this reduction type until much later. it should be noted at this time, due to the length of this paper, the comparisons made amongst the germanic languages will not be representative of a complete grammatical characterization. mainly, a brief description of nominal and pronominal markings will be discussed. in gothic, the original nominative singular ending of masculine a-stem nouns in proto-germanic was *-az. of all the germanic languages, gothic has remained closest to this, with its suffix –s (20): (20) goth. ohg dags ‘day’ tag (walters) the gothic nominative plural of the same class has the ending –ôs, which significantly differentiates gothic from some—but not all—other germanic languages (21): (21) goth ogh fuglôs ‘birds’ fogala (walters) the third person singular masculine personal pronoun in gothic is is. this form differentiates gothic from a number of languages in which that pronoun begins with h (22): (22) goth oe is hē (walters) 11 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 unlike some other germanic languages, gothic regularly distinguishes between the accusative and dative cases in the first and second person singular pronouns (23): (23) goth oe acc mik mē ‘me’ dat mis mē ‘me’ acc þuk ðē ‘thee’ dat þus ðē ‘thee’ (walters) contrasted against gothic, old norse (on) preserves the ending *-az of proto germanic in the nominative singular both of masculine a-stem nouns and of most strong masculine adjectives as –r, by way of runic –ar(24): (24) on goth ohg armr arms arm ‘arm’ góðr gôþs guot ‘good’ (walters) the nominative plural of the same masculine a-stems (although not of the adjectives) is expressed by means of the suffix –ar (25): (25) on goth ohg armar armôs arma ‘arms’ fuglar fuglôs fogala ‘birds’ (walters) in the masculine and feminine third person personal pronouns, on shows forms beginning in h-, unlike several of the other languages, including gothic (26): (26) on goth hann is ‘he’ honum imma ‘him’ (dat. sg.) hon si ‘she’ hennar izôs ‘her’ (gen.sg.) (walters) like gothic, but unlike a number of the other languages, on normally shows a distinction between accusative and dative in the first and second person singular personal pronouns (27): (27) on oe acc mik mē ‘me’ dat mér mē ‘me’ 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 13 acc þik ðē ‘thee’ dat þér ðē ‘thee’ (walters) old frisian shows no ending for the nominative singular masculine a-stem nouns, nor for the nominative singular of masculine strong adjectives (28): (28) of goth wei wigs ‘way’ gôd gôþs ‘good’ (walters) the nominative plural ending of the masculine a-stem nouns is variable in of, alternative between –ar or –er, although sometimes with –a. however, in some of the western dialects, it is -a and –an or –en, rather than –ar or –er. and as in oe, of has third person personal pronouns beginning with hthroughout (29): (29) of goth hi is ‘he’ him imma ‘him’ (dat. sg.) hiu si ‘she’ hire izôs ‘her’ (gen. sg.) hit ita ‘it’ (walters) like of, old saxon masculine nominative singular ending of both a-stem nouns and strong adjectives disappears completely (30), contrasting sharply with gothic: (30) os goth of dag dags dei ‘day’ gôd gôþs gôd ‘good’ (walters) the nominative plural of the masculine a-stem nouns in os is –os (31): (31) os goth ohg fuglos fuglôs fogala ‘birds’ (walters) the masculine third person personal pronoun in os shows forms beginning with hin the nominative singular, with much less frequent occurrence in such forms in other cases. although this feature distinguishes os clearly from gothic, the saxon forms are also different from on, which shows much more widespread use of h(32): (32) os goth on hê is hann ‘he’ 13 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 14 imu imma honum ‘him’ (dat. sg.) siu si hon ‘she’ ira izôs hennar ‘her’ (gen. sg.) (walters) additionally, most os texts do not distinguish between accusative and dative in the first and second person singular personal pronouns (33): (33) os goth acc mî mik ‘me’ dat mî mis ‘me’ acc thî þuk ‘thee’ dat thî þus ‘thee’ (walters) again, contrasting with gothic, old low franconian, in the nominative singular of masculine a-stem nouns, shows no ending (34): (34) olf goth day dags ‘day’ (walters) (the same holds true for the masculine nominative singular of strong adjectives.) all undisputed masculine a-stem nominative plurals show the ending –a (35): (35) olf goth daga dagôs ‘days’ (walters) the only personal pronoun in olf that shows an initial his the masculine nominative singular (36): (36) olf goth he is ‘he’ hie imma ‘him’ (dat. sg.) (walters) and as in os, there is no distinction between accusative and dative in the first and second person singular personal pronouns in olf (37): (37) olf os goth acc mi mî mik ‘me’ dat mi mî mis ‘me’ acc thi thî þuk ‘thee’ 14 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 15 dat thi thî þuk ‘thee’ (walters) and finally, old high german is also contrasted against gothic because ohg shows no trace of the original *-az ending of the nominative singular in the masculine a-stem nouns (38): (38) ohg goth tag dags ‘day’ (walters) ohg has also lost the same ending in the masculine nominative singular of strong adjectives. there a new ending –êr is frequently found, which has been added by analogy to the demonstrative pronouns: blint or blintêr ‘blind’ (walters). the nominative plural of the masculine a-stem nouns in ohg is regularly –a (39): (39) ohg goth berga bergôs ‘mountains’ fugala fuglôs ‘birds’ (walters) in the third person personal pronouns, ohg in general diverges sharply from the other germanic languages by having no forms in h(40): (40) ohg on of ër hann hi ‘he’ sīn honum him ‘him’ (dat. sg.) siu hon hiu ‘she’ ira hennar hire ‘her’ (gen. sg.) iz hinn hit ‘it’ (walters) ohg harshly diverges once again when dealing with the accusative and dative of the first and second person singular personal pronouns (41): (41) ohg oe olf os goth acc mih mē mi mî mik ‘me’ dat mir mē mi mî mik ‘me’ acc dih ðē thi thî þuk ‘thee’ dat dir ðē thi thî þuk ‘thee’ (walters) 15 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 16 4. the consequences of reduction and why english changed despite a common belief to the contrary, the loss of case-marking distinctions was a surprisingly orderly and systematic affair, at least in the south of england where substantial records from the period document the disappearing distinctions. although many treatments of case marking in me have given the impression of widespread confusion, the picture which emerges from a systematic study of individual texts is one particular form encroaching on the territory of other forms while category distinctions remain pretty much intact despite the existence of widespread syncretism. these findings go against at least certain variants of the hypothesis that the simplifications of morphology which occurred in me were due to creolization; for example, they are inconsistent with the idea that deterioration of the case-marking system of english was mainly due to incomplete language-learning on the part of french speaker who failed to master the case-marking system of english. a number of writers on oe have suggested that inflections became largely nonfunctional as a result of the growing anglo-norse contact: as anglo-norse contact grew, the case-marking system in oe atrophied and were lost. from this loss, the concomitant development of fixed word-order resulted. bradley’s (1904) discussion of the issue is old and informal, but still seems to be highly lucid and full of good sense: let it be imagined that an island inhabited by people speaking a highly inflected language receives a large accession of foreigners to its population. […] in our imaginary island, the foreigners will soon pick up a stock of words; if the island language is like the germanic ones, in which the main stress is never on the inflexional [sic] syllables, their task will be much easier. the grammatical endings will be learnt more slowly, and only the most striking will be learnt at all. the natives will soon manage to understand the broken jargon of the new comers, and to adopt it in conversation with them, avoiding the use of those inflexions which they discover to be puzzling to their hearers. but if they acquire the habit of using a simplified grammar in their dealings with foreigners, they will not entirely escape using it in their intercourse with each other. if there is intermarriage and absorption of the strangers in the native population, the language of the island must in a few generations be deprived of a considerable number of inflexional [sic] forms. let us now consider a somewhat different case. suppose that the two peoples who live together and blend into one, instead of speaking widely distinct languages, speak dialects not too far apart to allow of a good deal of mutual understanding from the first, or at any rate as soon as the ear has been accustomed to the constant differences of pronunciation. the two dialects, let us suppose, have a large common vocabulary, with marked differences in inflexion—a very frequent case, because phonetic change is apt to cause greater divergences in the unstressed endings than in the stressed stems of words. the result will be much the same as when peoples speaking distinct languages are mingled; indeed there are reasons for thinking that the change will be even more rapid and decisive. for one thing, the blending of the two peoples is likely to 16 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 17 take place more quickly. then, as the speakers of neither dialect will be disposed to take the other as their model of correct speech, two different sets of inflexional [sic] forms will for a time be current in the same district, and there will arise a hesitation and uncertainty about the grammatical endings that will tend to render them indistinct in pronunciation, and hence not with preserving. (26-28) however, it should be noted that anglo-norse contact did not trigger the developments in oe, but merely augmented or accelerated existing tendencies, themselves largely a consequence of the germanic fixing of stress on the first syllable. with each invasion, oe changed more, similar to the hypothetical accounting given by bradley (1904: 26-28). and due to these changes, oe systematically evolved into a language that did not need a case-marking system to differentiate its constituents. these lost inflections resulted in a more stabilized syntax.14 although a detailed discussion comparing oe and its germanic cousins could not be included (dates of change in particular), due to length, it should be obvious that oe changed significantly while its cousins did not necessarily change, at least not at the same rate that oe changed. in order to do such a detailed study, much more research—as well as length—would be required. it has been hypothesized that oe changed due to the number of invasions that occurred. with each invasion came a difference in language, resulting in changes of the ‘native’ language. and as the invaders’ languages might not have been of germanic origin, the systems present within proto-germanic quite possibly could not have been maintained. 14. for further discussion on the development of a stable word order, see bean 1983. 17 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 18 abbreviations acc = accusative dat = dative eme = early middle english gen = genitive goth = gothic instr = instrumentive masc = masculine me = middle english mode = modern english nom = nominative oe = old english of = old frisian ohg = old high german olf = old low franconian os = old saxon pl = plural sg = singular un = unmarked ælc.p = homilies of ælfric: a supplementary collection. (see references) cited by homily and line number. ælc.th = the homilies of the anglo-saxon church. (see references) cited by volume, page and line number. aw = the english text of the ancrene riwle: ancrene wisse. (see references) cited by page and line number. pc = the peterborough chronicle 1070-1154. (see references) 18 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 19 appendix a table 3. oe general masculine declension sg. pl. nom. se cyning ‘the king’ þā cyningas acc. þone cyning þā cyningas gen. þæs cyninges þāra cyninga dat. instr þæm, þÿ cyninge þæm cyningum table 4. oe general neutral declension sg. pl. nom. acc. þæt scip ‘the ship’ þā scipu gen. þæs scipes þāra scipa dat. instr þæm, þÿ scipe þæm scipum table 5. oe general female declension sg. pl. 1. nom. sēō talu ‘the tale’ þā tala acc. þā tale þā tala gen. þære tale þāra tala dat. instr. þære tale þæm talum 2. nom. sēō glōf ‘the glove’ þā glōfa acc. þā glōfe þā glōfa gen. þære glōfe þāra glōfa dat. instr. þære glōfe þæm glōfum table 6. oe the –an declension masc. fem. neut. sg.nom. se guma ‘the man’ sēō byrne ‘the coat of mail’ þæt ēāge ‘the eye’ acc. þone guman þā byrnan þæt ēāge gen. þæs guman þære byrnan þæs ēāgan dat. instr þæm, þÿ guman þære byrnan þæm, þÿ ēāgan pl.nom.acc. þā guman þā byrnan þā ēāgan gen. þāra gumena þāra byrnena þāra ēāgena dat. instr. þæm gumum þæm byrnum þæm ēāgum note: there are going to be exceptions, as with every language, but in the interest of length, the general paradigms are the only ones discussed and/or illustrated 19 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 20 appendix b table 7. declension of adjectives in oe masculine feminine neuter a.‘strong’ adjectives* example: til ‘good’ singular—all genders nom til tilu til acc tilne tile til gen tiles tilre tiles dat tilum tilre tilum plural—all genders nom tile tila tilu acc tile tila tilu gen tilra tilra tilra dat tilum tilum tilum b. ‘weak adjectives** nom tila tile tile acc tilan tilan tile gen tilan tilan tilan dat tilan tilan tilan plural—all genders nom tilan acc tilan gen tilra, -ena dat tilum *if stem is long, the –u feminine nominative singular and the neuter nominative/accusative plural is deleted, as in gōd ‘good.’ **very roughly, the ‘strong’ form of an adjective was used when the adjective was not preceded by a determiner, and the ‘weak’ form was used when a determiner was present. 20 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 loss and consequence 21 table 8. paradigm of the definite determiner in oe. masculine feminine neuter singular—all genders nom se sēo þæt acc þone þā þæt gen þæs þæ”re þæs dat þæ”m þæ”re þæm plural—all genders nom þā acc þā gen þāra dat þæ”m 21 walters: loss and consequence published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 22 references allen, c. 1995. case marking and reanalysis: grammatical relations from old to early modern english. oxford: clarendon press. allen, c. 1986a. “dummy subjects and the verb-second ‘target’ in old english.” english studies 6: 465-470. andrews, a. 1985. “the major functions of the noun phrase.” in timothy shopen (ed.) language typology and syntactic description, 62-154. cambridge: cambridge university press. bean, m. 1983. the development of word order patterns in old english. london: croom helm. bradley, h. 1904. the making of english. london: macmillan. campbell, a. 1959. old english grammar. oxford: clarendon press. gneuss, h. 1996. language and history in early england. brookfield, vt. haiman, j. 1974. targets and syntactic change. the hague: mouton. healey, a. and r. venezky. 1980. a microfiche concordance to old english. toronto: centre for medieval studies, university of toronto. helfenstein, j. 1870. a comparative grammar of the teutonic languages. london: macmillan and co. hopper, p. and s. thompson. 1980. “transitivity in grammar and discourse.” language 56: 251-299. koopman, h. and d. sportiche. 1990. “the position of subjects.” lingua 85: 211258. lightfoot, d. (ed.). 2002. syntactic effects of morphological change. oxford: oxford university press. milroy, j. 1997. “internal vs. external motivations for linguistic change.” multilingua 16: 311-323. moore, s. 1928. “earliest morphological changes in middle english.” language 4: 238-266. 22 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/7 doi: https://doi.org/10.25810/k43m-zz07 colorado research in linguistics 6-2004 loss and consequence: an examination of the old english case marking system as opposed to that of other old germanic languages denise e. walters recommended citation against optional wh-movement practice and domination: toward a theory of political micro-economy colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by chad nilep. practice and domination: toward a theory of political micro-economy chad nilep university of colorado at boulder older siblings play a role in their younger siblings’ language socialization by ratifying or rejecting linguistic behavior. in addition, older siblings may engage in a struggle to maintain their dominant position in the family hierarchy. this struggle is seen through the lens of language and political economy as a struggle for symbolic capital. bilingual adolescent sibling interactions are analyzed as both acts of identity and expressions of symbolic power. this paper draws a theory of political micro-economy, which relates face-to-face interaction to larger structures of political economy through a process of fractal recursivity. 1. introduction the present work will seek to describe the behavior of four individual members of a family as both acts of identity and expressions of symbolic power. the family observed, mother mami, teenaged daughters otoe, 18, and yumi, 14, and ten year old son ryu, are japanese-americans currently living in colorado. mami is an issei or first-generation japanese american; her children are nisei second generation. the behavior of primary interest is the choice of code – either japanese or english – which each family member uses in particular situations. i will argue that these four individuals relate to one another based on hierarchical rank. moreover, there is evidence for multiple hierarchies or social arrangements, based on multiple roles or identities of each individual. individuals must thus negotiate their relative status (and solidarity) based on the relevant hierarchy for a particular activity or frame. in addition, i will argue that a position within a hierarchy is a social good, related to social power. that is, getting and holding a position within a hierarchy both requires and bestows power and privilege. the hierarchical structure of the family distributes power unequally among family members. thus, the subject positions of individual actors both emerge from this structure and reproduce it. the confluence of these factors echoes ortner’s (1989) expanded practice theory1, as well as symbolic and linguistic approaches to political economy (gal 1989, friedrich 1989, inter alia). in linguistic anthropology, perhaps the primary statement of political economy originates with sociologist pierre bourdieu. his ce que parler veut dire, published in 1982, was originally delivered to the association of french teachers in limoges in 1977. by the time the english translation, language and symbolic power, was published in 1. for ortner, practice theory is approached via structure, actor, and practice, as well as one term not treated here: history. it is in approaches to history that ortner suggests that practice theory parts ways with political economy. perhaps, then, it is not surprising that this work lacks a theory of history. 1 nilep: practice and domination published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 1991, bourdieu’s concepts of symbolic domination and the linguistic marketplace were already being applied by linguistic anthropologists (e.g. friedrich 1989, irvine [1989] 1997). although some linguists have approached the notion of symbolic domination and the linguistic market as a metaphor, language and political economy may be better described as a metonym – reference to an institution by properties of the institution. the institution here is unequal distribution of power, control, and autonomy, as well as capital; the property by which it is referenced is the exchange of goods, the market. what approaches to political economy in the fields of linguistic and cultural anthropology have in common with economics as a field is an interest in unintended consequences. as the individual capitalist is apt to maintain the wealth of the state as though guided by an invisible hand (smith 1993), so the individual speaker/hearer is liable to maintain the hegemony of her culture or society in spite of her intent. javanaud makes clear the unintended consequences of cultural practices in his 1987 review of bourdieu’s ce que parler veut dire. bourdieu is not primarily dealing with conscious intentions. what governs groups of people (including philosophers) is not of a mechanical, cybernetical or system-theoretic nature. they act neither deterministically nor in [a] transparent teleological way. decisions may appear free and yet are constrained. our decisions are not normally the result of cynical calculations, even if such calculations can enter into them. rather, our habitus implies a disposition to act conditioned by its acquisition and use on certain markets (810-11). the behavior analyzed here can be seen as acts of identity (romaine 1988). according to bucholtz and hall (2003), the study of linguistic anthropology is largely the study of language and identity. i seek here to relate notions of political economy with notions of personal or local identity2. this study entails a concern for the production of individual subject positions, including the means by which identities are created, socialized, and reproduced. i argue that older siblings serve to socialize younger ones into locally appropriate roles, while at the same time working to build and maintain their own roles. this identitywork yields symbolic capital to the older sibling and creates the market for its exploitation by constructing appropriate social hierarchies. the notion of symbolic capital comes, of course, from studies of language and political economy. such studies originated with marxist scholars’ turn toward symbolic anthropology and discourse analysis, as well as linguistic anthropologists’ turn toward power and institutions (gal 1989). as such, studies of political economy often take as their analytic focus nations (e.g. irvine 1997), or major class or language divisions within a state (e.g. bourdieu 1977, woolard 1989). however, the ‘common sense’ of the nation – if there is such a thing3 – must be constructed through the common senses of its members. the collective identity and collected practices that help constitute the nation are experienced at the individual level, by actors in face-to-face contact with other individuals (suleiman 2003). 2. see walters (1996) on the relations of language, identity, gender, and political economy. 3. herzfeld (2001), for example, suggests that nationalism or ‘national culture’ conflates cultural boundaries with the borders of the state. 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/8 doi: https://doi.org/10.25810/gx41-kq64 practice and domination: toward a theory of political micro-economy 3 the link between the individual experience necessary for socialization and the collective experience of the linguistic market is made via fractal recursivity (irvine and gal 2000). in mathematics, a fractal is a rough geometric shape that is self-similar at different scales. in other words, smaller portions of the form resemble larger portions, as well as the form as a whole (at least approximately). this notion of fractal geometry is useful in social science4, where individual patterns of behavior or thought can be seen to recur in larger group behavior. irvine and gal identify fractal recursivity, together with iconization and erasure, as semiotic processes by which individuals construct ideological representations of distinction5. perceptions of opposition at one level of analysis are projected onto other levels, recreating the ideology of distinction. for example, opinions about an individual are attributed to a group the individual belongs to. fractal recursivity is not only useful for describing the individual’s ideology of distinction at various levels of analysis, but also for relating individual behavior to the structure of society. sapir ([1927] 1995) suggested that the difference between individual and society is largely a matter of the analyst’s focus. society is thus seen as the collected senses of individuals. particular differences do not obliterate the self-similarity of the individual and the group; like any fractal, form is independent of scale. thus, i assume that an extremely local analysis is a proper starting point for the analysis of ‘culture’ or ‘common sense’ (herzfeld 2001). a microanalysis of the individual members of mami’s family should illuminate notions of symbolic domination, socialization, and practice at other, more macro-levels. this is not to say that the analysis of family interaction is the only site at which to study these notions. indeed, the insensibility of a culture from the outside suggests that analysis must ultimately look more widely for the borders of common sense. however, i hope that this analysis might provide a preliminary step to mediate between the experience of face-to-face interaction and macro-historical processes and the exercise of institutional power (gal 1989: 34950). 2. data what follows is a close reading of interactions between mami and her family. transcripts of audio recordings are included to illustrate the specific actions of each individual, and the way they construct and reflect social role. in addition to the linguistic data illustrated here, analyses are based on field notes made during participant observation of various japanese american families in the colorado front range. however, the primary focus is on code switching and code choice as a particularly visible site for the assessment of social role. in the data that follow, mami speaks japanese to her children. mami is a native speaker of japanese, with some english proficiency. all three children are balanced, 4. irvine and gal (2000) trace the notion of fractals in anthropology as far back as bateson (1936). for discussion of fractal geometry in economics, see takayasu et al. (2000). 5. roughly, iconization is the process whereby indexical markers of a social group are transformed into iconic representations of the group, ‘naturalizing’ perceptions of the other. erasure is the discounting or ignoring of facts inconsistent with an ideological scheme, so that facts inconsistent with ideology do not count as counter-evidence. 3 nilep: practice and domination published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 simultaneous bilingual speakers6 of english and japanese. within the home, mami prefers to use japanese as a means of heritage language maintenance (c.f. cashman 2001, langager 2001). she requires that her children speak japanese at home, with the understanding that they will rely primarily on english outside the home. although i gathered no data from outside the home, it seems safe to assume that very little interaction outside the family takes place through the medium of japanese7. according to an analysis of the 2000 us census carried out by the social science data analysis network (2002), approximately 85% of colorado residents are english monolinguals. in addition, approximately 11% of colorado residents speak spanish, while fewer than two percent speak an “asian language.”8 thus, the japanese speaking population of the region may be assumed to be quite small. 3. analysis since mami speaks japanese to her children, the language serves an indexical function (silverstein 1995). mami’s use of japanese to enact behaviors such as teaching, directing action, and scolding builds a semiotic link between these acts and the language used to perform them. thus, japanese indexes social roles such as teacher, director, and authority. in excerpt 19, mami both indexes and achieves her role as mother through appropriate actions, which yumi ratifies in turn. according to ochs (1993), identity should not be understood as a static property of the individual, but rather as a position that must be constantly achieved and ratified through interaction. yumi similarly achieves her identity as daughter through locally appropriate behavior. (1) excerpt 1 ‘sansuu no mondai’ 1 m: [ni-juu kyu doru kyu juu ���������� � ��� ��� � ������ ������� ����� kyu sento desu 2 y: ee? � � � � 3 (2.0) 4 y: ni doru ha? � ��� ��� ��� 5 m: ha? ni-hyaku go juu � � � � �� �������� � � �� � � ����� 6 minutes de � ��� ��� 7 y: nn � � � 8 m: ni-juu doru kyu-juu kyu � ���� � ��� ��� � ������ ������� ����� 9 sento 6. these claims are made without the benefit of testing or serious inquiry into the nature of the children’s linguistic competence. however, as the data show, each child is able to produce fluent discourse, both formally and pragmatically consistent with native-speaker competence in english as well as japanese. this paper will not address psycholinguistic questions of the nature of bilingual competence. 7. one exception is the japanese-medium supplementary school that ryu attends on saturdays. 8. ssdan defines “asian language” as any non-indo-european language spoken in asia. 9. throughout the paper, japanese data is presented in hepburn romaji. free idiomatic translation is included in the right column. 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/8 doi: https://doi.org/10.25810/gx41-kq64 practice and domination: toward a theory of political micro-economy 5 10 y: socchi no hoo ga yasui jan � � ���������� � � � ��� 11 m: ima . nana-juu go � ���������� ��� ���� ��� ��� ��� 12 minutes de . juu kyu ��������� ��� ��� � ������ ������� ����� 13 doru kyu-juu kyu sento desu 14 (m): [(inaudible) 15 y: [(inaudible) hyaku �� ������� � � �� �� ��� ������� 16 minutes no wa? 17 m: konkai hyaku minutes � � ������ ����������� � � �� �� ��� ��� 18 (tsukaimashita) ������� �� �������� � �� �� ���� ��� ����� ��� ��� 19 san-juu ichi doru 20 kudasareba .. 21 docchi ga ii desu ka? � �� � ��������! ������ 22 hai sansuu no mondai " �� � �������� �� �� �� ��! ��� � 23 y: mukoo no yatsu (tte iu ka) � � ������ 24 kaeta yatsu � � �������� ��� � �� � � 25 m: nn (soo desu) ne? " �� � � 26 y: yokatta ne. ne kore � � ����� �� ����� � ���� � ������ 27 wa [(dore) 28 m: [demo san # � ���� ���� 29 m: ni-juu kyu doru plus tax ����� ������ ��� ����� �� ��� $ ������� �� � � 30 ga hairu n yo 31 y: n � � � 32 m: ii yo ne. sore demo % �� ����� � ���& ��� ���� ��������������� � � � ��� 33 yasui kara ne in excerpt 1, mami shows (achieves) her position as mother by enacting several stances, which make up the role. at the same time, yumi enacts stances proper to her role as daughter. mami and yumi are discussing the price of long-distance calling cards. throughout the conversation, mami possesses information, which she gives to yumi. she repeatedly (lines 1, 5-7, 10-11, 14-15) gives yumi specific information about the cost of various cards and their denomination in minutes. then, at lines 16-17, mami asks yumi to decide which card is the best value, posing the question as an academic exercise. docchi ga ii desu ka? hai, sansuu no mondai. “which is better? right, it’s a math problem.” these acts position mami as teacher in this interaction. yumi’s responses, in turn, serve to ratify mami’s position. she evaluates mami’s information (line 8), asks for more information (line 13), and generally allows mami to maintain primary speakership (jefferson 1984). notice how the women use questions to accomplish different roles within the interaction. yumi’s question hyaku minutes no wa? “what about the one with a hundred minutes?” requests information that mami possesses and yumi wishes to 5 nilep: practice and domination published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 know, namely, the price of the phone card. however, mami’s docchi ga ii desu ka? “which one is better?” directs yumi to manipulate the information mami has provided her. this is not a request for information, since mami already possesses all of the relevant information and has just given it to yumi. instead, it is a directive, specifically marked as sansuu no mondai “an arithmetic problem.” by acting both as teacher and as director, offering information and directing yumi to manipulate it, mami enacts a position as mother. further, as the data show, mother is a relatively more powerful role: she is able to direct yumi’s action (lines 1617) and to evaluate her solution (line 21). furthermore, she is not obliged to respond to yumi’s questions or requests for clarification (lines 2-4 and 23). japanese and english are each associated with different social roles, an association built through a process of indexicalization. for example, though mami can speak english (for example, she speaks english to me), japanese is her preferred, dominant language. when speaking to her children, she almost always uses japanese. thus, since japanese is associated with a particular role within the family hierarchy, the language becomes an index of the role. further, using another language facilitates stepping outside that role to enact other aspects of personal identity. the use of english calls upon an alternate set of identities and relationships, an alternative identity market. excerpt two illustrates such a use of english. (2) excerpt 2. ‘lucky? happy?’ 1 y: doo shiyoo ka na � ���� �� � ��� �� 2 m: otoe-neechan ni mo aji ' �( ��������) �������� ���� ��������! ��� 3 chotto mite moratte 4 y: (inaudible) 5 m: [ha � � 6 o: [heh � � � 7 y: (inaudible) 8 m: lucky? happy? * � � ( � �+ � � � notice that the english used here is limited to single words. mami typically uses single words or short phrases when speaking to her children in english10. moreover, the words used appear somewhat ill suited. mami and her daughters are preparing dinner; yumi is making a salad dressing. mami instructs yumi to let otoe taste the dressing, and asks otoe for her opinion. while we might expect the dressing to be described as oishii “tasty,” mami uses english words that express a positive connotation, but would not be expected for describing food. this may suggest a general lack of english fluency. as we have seen, mami uses japanese when instructing yumi or giving her direct commands, enacting a role as teacher or director, consistent with her position as mother in the family hierarchy. however, she uses english words to elicit otoe’s opinion of the food offered. this change in language can index a change in role. mami is soliciting 10. similarly, when mami speaks to me, she typically uses short phrases or simple sentences. further, her most common speech to me is in the form of japanese minimal responses such as un “all right.” unfortunately, there are no recordings or careful field notes that reflect my conversations with mami. 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/8 doi: https://doi.org/10.25810/gx41-kq64 practice and domination: toward a theory of political micro-economy 7 otoe’s assessment, and marks the change with a change in code. the adjustment of roles allows the women to suspend the requirements of the family hierarchy, permitting otoe to serve as arbiter. in the next excerpt, mami instructs her daughters in the proper methods of skewering meat in order to cook it. two activities, cooking and teaching, are both associated with the role of mother, and thus are enacted through the medium of japanese. however, mami’s uncertainty, visible through the use of interrogative hedges (nan tte iu; dokka) and softeners, weakens her position here. while otoe continues to use japanese, maintaining the established frame, yumi breaks into english. i think they went diagonally when they pierced it. yumi does not receive instruction from mami, but offers it, inverting the teacher/student relationship. at the same time, the switch from japanese to english may locate this exchange outside the frame of the family hierarchy. (3) excerpt 3. ‘iie, nihongo’ 1 m: nan tte iu no dokka ni � ����� ������ �� � ����( ������� ��� � ����� �� 2 niku wo sashiten no ana �� ��� � � ��� ��� ��� 3 ana [ga aiteru toko] 4 o: [h niku] � � � 5 m: ni kooshite ne �� ���� � 6 o: datte= ���� 7 y: =i think they went ���� ��( ��� � ������ � � �� �� ���� �� 8 diagonally .. when they �� � �� ���� � ���� 9 pierced it 10 (1.4) 11 o: honma… hhh �� �� 12 y: i showed otoe my cool ���� ��� �) ����� �� ����! �� ������� 13 bath towel 14 o: [mmm] � � � 15 y: [don’t you] remember. was , ����� �� ���� �� ! ��� � ���������� ����� � 16 it on this [sside 17 m: [iie nihongo � ��� � ���� 18 o: [hh 19 y: [hh 20 o: dochi demo ii . ��� ������) / � 21 y: katte ni deta no ���0� ���� � ���� �� yumi’s english contribution is greeted by a long silence. otoe is the first to break this silence, speaking a single word11 of japanese at low volume. thus, otoe has not 11. otoe’s honma “really” is perhaps strategically ambiguous. it could be taken as either agreement with, or questioning of yumi’s assertion. 7 nilep: practice and domination published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 followed yumi out of the established frame. she does not take up yumi’s code choice, nor does she continue the discussion. following this problematic exchange, yumi changes the subject. her remark, i showed otoe my cool bath towel, is addressed to mami in english. while the use of english marks a break from mami’s lesson, it continues the code choice made in the problematic section. mami explicitly addresses this utterance as delivered in the wrong language: iie, nihongo “no, japanese.” by censuring talk, particularly the form of talk, mami presents herself as authority, attempting to reestablish herself at the top of the hierarchy. rather than acquiescing to the attempted censure, however, the daughters take up mami’s utterance as the opening of a discussion. yumi’s plea, katte ni deta no “it just came out,” may be taken as an appeal addressed to an authority. however, otoe does not recognize mami as authority, instead placing herself in that role and judging yumi’s contribution as in bounds. dochi demo ii “either way is ok.” otoe often places herself in the role of authority, as in the preceding exchange. this strategy puts her in a locally strong position relative to those being judged. this can test mami’s dominant position, as in excerpt three, and therefore challenge the family hierarchy. more often, though, otoe serves as authority in disputes among her siblings, as in the following passage. this is a proper role within the family frame, given otoe’s role as older sibling. (4) excerpt 4. ‘excuse my pleasure’ 1 r: ando. � 2 (1.8) 3 y: (dakara) ni man & �������� ���� ���� �� � � 4 [(yon sen) 5 o:[ni man go sen tte. itta 1��2�� � ������ ���� ���� �� � � ������������ ���� � 6 jan .. ni man go sen �� �� � � � 7 (2.7) 8 y: ˚(it worked before)˚ .. ������( � �! ������ 9 o: hh 10 (1.5) 11 y: it’s been my: pleasure �����! ����� �� �� �� ��� ((�������� ��� ���))12 12 o: h[hh 13 r: [hh 14 y: hai. go sen mo " �� � ������ ���� �� � � �� ��� 15 [(inaudible) 16 o:[no no no no no ���������� 17 r:[hhh 12. transcriber’s notes in lines 11 and 18 give a phonetic transcription in order to illustrate phonotactic differences. 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/8 doi: https://doi.org/10.25810/gx41-kq64 practice and domination: toward a theory of political micro-economy 9 18 o: excu:se my: pleasure 3. $ � � ���� �� �� �� ��4 (( �� ��� ���� ����� ����� �� )) 19 r: hh ((claps hands)) 20 o: hai � � 21 o: soo soo ��� � ������ � � 22 r: a arf arf ��� �� 23 o: aho. aho. � ������ ��� here, otoe, yumi, and ryu are playing a board game. it seems that yumi is trying to cheat, offering to pay twenty-four thousand dollars (line 3), when she owes twenty-five thousand. otoe corrects her, and yumi gathers the correct sum. at line seven, yumi switches from japanese to english, offering a sotto voce aside in which she reveals that she may be in the habit of misrepresenting the sums she owes: it worked before. this switch from game play to evaluation is a change in footing (goffman 1979), and thus an appropriate occasion for code switching. however, when yumi returns to game play (line 10), her continued use of english is seen as a problem. at line 10, yumi adopts a very affected tone of voice as she hands over the money she owes. she speaks in a lower pitch than usual, and stretches her words, perhaps attempting to play some sort of character. moreover, she speaks english, and uses a somewhat stilted expression. her odd performance is greeted by laughter from otoe, with ryu joining in immediately after (lines 11-12). yumi attempts to escape this laughter by returning straight away to the game. otoe13 will not allow yumi to elude judgment, however. at line sixteen, she parodies yumi’s odd expression. rather than simply repeating what yumi said, however, otoe reworks the expression slightly, rendering it utterly nonsensical. excuse my pleasure. moreover, although otoe, like yumi, uses a low pitch, she delivers the expression with an exaggerated japanese accent14. in this interaction, otoe is able to present herself as the linguistic authority. she treats some code switching as appropriate and acceptable, as when yumi (line 7) or ryu (line 1) evaluate or comment on the game. similarly, otoe herself sometimes uses code switching to show such changes in stance or topic. this is similar to the use of code switching for realignment or control that zentella (1982, 1997) describes among spanishenglish bilinguals. however, not all code choices are treated as legitimate. otoe judges her siblings’ language behavior, censuring inappropriate conduct. by selectively ratifying certain actions, but censuring others, otoe helps to teach yumi and ryu what range of behavior is considered appropriate within the family15. at the same time, though, otoe claims for herself the same powerful position that mami serves in other interactions. as with all such practice, the effects of the interaction have a range of consequences for each participant. 13. still problematic for this analysis is otoe’s “no, no, no” at line 16. if, as i argue, otoe uses japanese as an index of power and authority, why is this initial censure delivered in english? 14. neither yumi nor otoe normally have a japanese accent when speaking english. 15. this type of sibling behavior may also serve as a bridge to outside communities (mannle and tomasello 1987). however, yumi and ryu are well beyond the age of children most often implicated in the bridge hypothesis. 9 nilep: practice and domination published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 as these excerpts illustrate, the members of mami’s family conduct themselves within a family hierarchy. their behavior toward one another is partially structured according to locally relevant norms of interaction. as mother, mami is entitled to judge the behavior of her children, and to direct their behavior. at the same time, she is obliged to offer them instruction and care. conversely, the children are not expected to offer mami instruction or judgment. that is not to say that the children do not offer these things, but that when such perturbations occur, they are marked as out of the ordinary. silences and code switching provide particularly visible marks of these shifts. mami sits atop a family hierarchy. the hierarchy is further ranked by age divisions, with older sister otoe holding more power than either of her siblings. this should not be taken to mean, however, that age itself is constitutive of position. rather, the hierarchy is created through social practices, each of which is subject to ratification by fellow actors. as excerpt five shows, even the youngest member of the family can claim the privileged position of authority by refusing to ratify invalid acts. (5) excerpt 5. ‘kyuu ryoobi’ 1 r: ˚ichi ni san shi go roku ) ��������� ������� ����� ����$ ���� ������ � ���������� 2 shichi hachi kyu to˚ 2 o: (inaudible) 3 m: soo ne � � ������� � �� 4 r: kyuu ryoobi 5 � � 5 o: hold on wha+ �� ������ � 6 you’re there already �� ������ ���� ��� � 7 r: ni man go sen (aru yo) (3) � ���� ���� ���� �� � � � 8 y: (shotokuzei mo) ' � ���� �� ��� $ � 9 r: (sore ichi-man yon-sen) � � ������� �������� �� � � � 10 o: thís is hé:rs � � ������� ���� 11 y: yeah � � 12 r: ni-man yon-sen � ���� ���� ����� �� � � � 13 o: nande? � � 14 r: nikai aru kara kyuu ryoobi ga # �� � ����� ���� �������� � �� 15 y: aa honto da ) � ���� ������� � �� excerpt five occurs just prior to excerpt four, above. recall that while yumi’s use of english within game play was ruled out of bounds, switching to english as a change in footing received no censure. again in excerpt five, english is used outside of actual game play. at line five, otoe challenges ryu, who is demanding to be paid for completing a round of the game. otoe switches to english, marking her role as authority16. ryu does not respond to this challenge, however. at line seven, he reinstates his demand by 16. most often otoe uses japanese, with its indexical power, to challenge or command her siblings. however, a code switch is warranted here as what zentella (1997) calls a control switch. 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/8 doi: https://doi.org/10.25810/gx41-kq64 practice and domination: toward a theory of political micro-economy 11 naming – in japanese – the figure he is to be paid. yumi follows ryu, both in code choice and by addressing herself to the game, not otoe’s criticism of ryu. otoe’s next challenge (line 10) is made more forcefully, again in english. this time, yumi follows otoe, using the english yeah to voice agreement. however, ryu refuses to recognize otoe’s censure, continuing to play the game and to speak japanese. by not ratifying otoe’s stance as authority, ryu has effectively denied her that stance. this behavior, with yumi’s eventual ratification (line 15), allows ryu to win the exchange, and illustrates the interactional nature of these stances. 4. discussion the family hierarchy alluded to here is not a structure that exists independently of the actions of family members. rather, it is a web at once constraining the actions of family members, and woven by them (geertz 1973). mami’s behavior serves to create and to index her own identity, to socialize her children into locally relevant patterns of ‘common sense,’ and to construct the hierarchy within which it has power and significance. this hierarchy of power and significance can be seen as a local symbolic market. the arrangement of subject positions with unequal access to approved forms of behavior, and unequal rights to exercise power bears a fractal relationship to the symbolic marketplace described by bourdieu (1977, 1991) and others. that the individuals examined here may be seen to constitute the very market in which they participate argues against bourdieu’s assertion that the state, in the form of the educational system, has sole control over the creation of symbolic markets. according to bourdieu, “the educational system is a crucial object of struggle because it has a monopoly over the production of the mass of producers and consumers, and hence over the reproduction of the market on which the value of linguistic competence depends” (1977: 652). although the educational system may be among the most common of institutions within many cultures, it does not have a monopoly over social actors17. each individual performs within a range of institutions and frames, with a number of fellow actors. each of these interactions can have a constitutive effect on not only the identity of the individuals, but also on market formation, distinction, and symbolic domination. the traditional, ‘macroeconomic’ view of political economy focuses on societies, regions, and nations. a microeconomic approach to symbolic domination takes seriously suleiman’s reminder that institutions such as the nation, as a conflation of the political state with a ‘culture,’ are comprised of individual actors. “[collective] identities are experienced at the personal level[;] it is the individual who experiences these identities and gives them meaning in his or her social and cultural setting” (suleiman 2003: 5). political micro-economy takes as its focus the daily interactions of individuals as both acts of identity and steps in the constitution of symbolic markets and individual habitus. what this study has not shown directly, what remains for future work, is the effect of face to face interaction on larger structures of society. however, it is hoped that the notions of fractal recursivity, self-similarity, and scale independence allow an avenue to 17. for fuller critiques, see woolard (1985), gal (1989), briggs and bauman (1992), and hill (1993), inter alia. 11 nilep: practice and domination published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 trace the personal practices of individual identity to social orders and schema constructions within the broader culture. we are often reminded that social positions such as gender (eckert and mcconnell-ginet 1992; ortner 1996) and ethnicity (said 1979, suleiman 2003) are the constructions of individual actors within an encompassing structure. as eckert and mcconnell-ginet put it, “it is the mutual engagement of human agents in a wide range of activities that creates, sustains, challenges, and sometimes changes society and its institutions” (462). this study has been an attempt to show the micro-level construction of shared cultural schemas. future work should take a gradually more encompassing view of social interaction among families, congregations, neighborhoods, cities, and even nations. such an expanding focus is necessary to trace the borders of particular symbolic economies. i expect that, within a macro-economy, standards of practice among larger groups will bear a fractal resemblance to smaller ones. 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/8 doi: https://doi.org/10.25810/gx41-kq64 practice and domination: toward a theory of political micro-economy 13 references bateson, gregory. 1936. naven, a survey of the problems suggested by a composite picture of the culture of a new guinea tribe drawn from three points of view. cambridge: the university press. bourdieu, pierre. 1977. outline of a theory of practice. cambridge: cambridge university press. bourdieu, pierre. 1982. ce que parler veut dire. paris: faynard. bourdieu, pierre. 1991. language and symbolic power. cambridge, ma: harvard university press. briggs, charles l. and richard bauman. 1992. ‘genre, intertextuality, and social power.’ journal of linguistic anthropology 2(2): 131-172. bucholtz, mary and kira hall. 2003. ‘language and identity.’ in alessandro duranti (ed.) a companion to linguistic anthropology. malden, ma: blackwell. cashman, holly. 2001. doing being bilingual: language maintenance, language shift, and conversational codeswitching in southwest detroit. ann arbor: university of michigan dissertation. eckert, penelope and sally mcconnell-ginet. 1992. ‘think practically and look locally: language and gender as community-based practice.’ annual review of anthropology 21: 461-90. friedrich, paul. 1989. ‘language, ideology, and political economy.’ american anthropologist 91(2): 295-312. gal, susan. 1989. ‘language and political economy.’ annual review of anthropology 18: 345-67. geertz, clifford. 1973. ‘thick description: toward an interpretive theory of culture.’ the interpretation of cultures: 309-23. new york: basic books. goffman, erving. 1979. ‘footing.’ semiotica 25: 1-29. herzfeld, michael. 2001. anthropology: theoretical practice in culture and society. malden, ma: blackwell. hill, jane. 1993. ‘structure and practice in language shift.’ in kenneth hyltenstam & ake viberg (eds.) progression and regression in language: sociocultural, neuropsychological, and linguistic perspectives, 68-93. cambridge: cambridge university press. irvine, judith. 1997. ‘when talk isn’t cheap: language and political economy.’ in donald brenneis and ronald macaulay (eds.) the matrix of language: contemporary linguistic anthropology, 258-83. boulder: westview press. irvine, judith and susan gal. 2000. ‘language ideology and linguistic differentiation.’ in paul kroskrity (ed.) regimes of language: ideologies, polities, and identities, 35-83. santa fe, nm: school of american research. javanaud, pierre. 1987. ‘what is language all about?’ journal of pragmatics 11: 799815. jefferson, gail. 1984. ‘notes on a systematic deployment of acknowledgement tokens “yeah” and “mm hm.”’ papers in linguistics 17(1-4): 197-216. langager, mark. 2001. sojourning with children: the japanese expatriate educational experience. cambridge, ma: harvard dissertation. 13 nilep: practice and domination published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 14 mannle, sara, and michael tomasello. 1987. ‘fathers, siblings, and the bridge hypothesis.’ in keith nelson and anne van kleeck (eds.) children’s language iv. new york: gardner press. ochs, elinor. 1993. ‘constructing social identity: a language socialization perspective.’ research on language and social interaction 26(3): 287-306. ortner, sherry. 1984. ‘theory in anthropology since the sixties.’ comparative studies in society and history 26(1): 372-411. ortner, sherry. 1989. high religion: a cultural and political history of sherpa buddhism. princeton: princeton university press. romaine, susan. 1988. pidgin and creole languages. new york: longman. said, edward. 1979. orientalism. new york: vintage books. sapir, edward. 1995. ‘the unconscious patterning of behavior in society.’ in ben blount (ed.) language, culture, society, 64-84. prospect heights, il: waveland press. smith, adam. 1993. an inquiry into the nature and causes of the wealth of nations. kathryn sutherland, ed. oxford: oxford university press. social science data analysis network. 2002. censusscope. http://www.censusscope.org/index.html. suleiman, yasir. 2003. the arabic language and national identity. washington: georgetown university press. takayasu, hideki, misako takayasu, mitsuhiro p. okazaki, and tokiko shimizu. 2000. ‘fractal properties in economics.’ in miroslav novak (ed.) paradigms of complexity: fractals and structures in the sciences, 243-58. singapore: world scientific publishing company. woolard, kathryn. 1985. ‘language variation and cultural hegemony: toward an integration of sociolinguistic and social theory.’ american ethnologist 12: 738-48. woolard, kathryn. 1989. double talk: bilingualism and the politics of ethnicity in catalonia. stanford: stanford university press. walters, keith. 1996. ‘gender, identity, and the political economy of language: anglophone wives in tunisia.’ language in society 25(4): 515-55. zentella, ana celia. 1982. ‘spanish and english in contact in the united states: the puerto rican experience.’ word 33 (1-2): 41-57. zentella, ana celia. 1997. growing up bilingual: puerto rican children in new york. malden, ma: blackwell. 14 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/8 doi: https://doi.org/10.25810/gx41-kq64 colorado research in linguistics 6-2004 practice and domination: toward a theory of political micro-economy chad nilep recommended citation microsoft word paper_nilep.doc memories of everyday life in communist bulgaria: negotiating identity in immigrant narratives colorado research in linguistics. june 2006. vol. 19. boulder: university of colorado. © 2006 by nadia kaneva. memories of everyday life in communist bulgaria: negotiating identity in immigrant narratives nadia kaneva university of colorado at boulder the fall of the berlin wall in 1989 marked the end of the communist era and the political and economic structures that supported it. yet, for many eastern europeans communism was not a monolithic “evil empire” but their “normal” way of life. this paper focuses on the narratives of bulgarian immigrants to the us about experiences that formed the fabric of everyday life in communist bulgaria. the informants in this study are not political immigrants. they came to the us after 1989 in pursuit of educational and career goals and claim to have had “average” lives in bulgaria. however, they belong to a generation that came of age in the last years of communist rule in bulgaria and have a unique perspective on that period. the analysis approaches memory and identity as narrative constructions that are constantly renewed, struggled over, and adapted to the present context. in exploring this instability, the paper seeks to identify common patterns among the stories told by immigrants, which represent pieces of the collective memory of ordinary life under communism in bulgaria. 1. introduction: what was communism? with the fall of the berlin wall in 1989, “actually existing communism” in central and eastern europe was pronounced dead although the reasons for its collapse are still the object of debate and research (e.g., verdery 1996, burawoy & verdery 1999). while ideological accounts of the soviet bloc often portrayed it as a monolith of oppression, communism as a social practice and lived experience did not have a single face in the different countries of central and eastern europe, nor did it have a fixed form during its existence over the course of half a century. the task of writing the history of communism is complicated by the emergence within public discourse of various personal accounts that had been previously suppressed by totalitarian regimes, the declassification of state archives, and the opening up of spaces for collective remembering and questioning of the past. perhaps the most controversial and painful memories to emerge in the process of reassessing the communist past are those of survivors of political oppression, among whom are survivors of various internment camps (e.g., ratushinskaya 1988, sherbakova 1992, todorov 1999). these personal “survivor narratives” have come, at least in western eyes, to represent everything that was horrible about communism and are often used to reaffirm preexisting stereotypes about the corrupt nature of the communist system. however, for many people in central and eastern europe life under communism was simply their “normal” way of life. indeed, it was no less filled 1 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 2 with human emotions and struggles than life under any other system, although the nature of these struggles was inevitably influenced by different socio-political, economic, and ideological conditions. in this paper, i focus on a set of narratives about experiences that were seen as mundane and formed the fabric of everyday life in communist bulgaria. i analyze the personal narratives of bulgarian immigrants to the us who emigrated after 1989. the informants in this study belong to a generation that came of age in the last years of communist rule in bulgaria and have a unique perspective on that period. they are not political émigrés or dissidents and do not see themselves as “communism survivors.” on the contrary, they claim to have had “average” lives in bulgaria and came to the us legally in pursuit of educational and career goals. examining their stories may contribute to the reconstruction of a fuller memory of “life under communism” – one that is not fixated on political oppression but allows for mundane and peaceful moments to be remembered and told. 2. theoretical focus: the intersection of memory and identity this analysis of the personal narratives of immigrants explores the intersection of memory and identity and is situated within a constructivist theoretical framework. from this vantage point, reality is viewed as socially constructed (berger & luckmann 1967) and may be understood as “a scarce resource” that is produced and contested through communication (carey 1989: 87). in this view, control over communication processes is central to the struggle over the nature of the real, and personal narratives are one arena where this struggle can be evidenced and explored. consistent with this framework is the work of scholars influenced by symbolic interactionism who see personal and collective identities as products of social interaction, which become present and known through narratives (goffman 1959, bruner 1991, gagnon 1992, holstein & gubrium 2000). within this view of identity, memory narratives become important acts of identity production and sites for the negotiation of meaning. in my analysis of memory, i rely on the theoretical work of maurice halbwachs (1980) who coined the term collective memory in his book la memoire collective, originally published in france in 1950. halbwachs establishes several central principles of collective remembering. first, collective memory is constructed through communication, and depends on the existence of an “affective community” which can sustain it through its communication practices (31). second, memory is always embedded in a spatial and temporal dimension (187). that is, we remember events by remembering specific places and moments that are linked together into a narrative. finally, memory narratives are always reconstructions, which serve purposes rooted in the present. thus, memory is unstable and often incorporates events in the present into the telling of the past (69). the three dimensions of collective memory identified by halbwachs – affective community, space, and time – are central to my analysis of immigrant 2 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/2 doi: https://doi.org/10.25810/tjqa-yy74 memories of everyday life in communist bulgaria 3 narratives. i attempt to understand how each of these dimensions is embedded in narrative form by bringing to bear the theoretical notions of frameworks (goffman 1974, tannen and wallat 1987) and narrative time (ricoeur 1980). in sum, i approach memory and identity as narrative constructions that are constantly renewed, struggled over, and adapted to the present context. in exploring this instability, i seek to identify common patterns among the stories told by immigrants, which represent pieces of the collective memory of ordinary life under communism. 3. methodology: meet the immigrants this study adopts an ethnographic approach and is based on a set of narratives by eight bulgarian immigrants to the us, which were collected over the course of four months between december and april 2004. all of the participants were living in the denver, colorado metro area at the time of the study. all of them knew each other and formed a loose network of friends and acquaintances. i met them at different times in the fall of 2003 and maintained casual contact with them for over a year before conducting the interviews for this study. my initial introduction to the group was not as a researcher but simply as another bulgarian living in the area. while i did not have the idea for this study at that time, i shared that my area of research was communication and that i was interested in studying bulgaria and communism. my interactions with the people in this study over the course of the year were usually associated with celebrations of birthdays, bulgarian holidays, or simply social visits. one of the joys of these contacts came, as stated by many in the group, from the opportunity to speak our native language and talk about topics that would be unfamiliar or strange to americans. a personal interest in crosscultural communication motivated me to observe which topics were deemed particularly “foreign” to americans. on several occasions i would hear the phrase, “how can you explain this to an american?” and noticed that often it referred to the inability to communicate a way of thinking or acting that was embedded in bulgarian culture, defined as “a whole way of life” (williams 1977). this last realization prompted me to conduct more focused conversations with several of the people in the group during which i asked them to recollect in more detail their life in bulgaria and talk specifically about how they recall “life under communism.” the excerpts presented in this paper come from two individual interviews and two group interviews with five and three participants respectively. a total of eight people were interviewed, including three women and five men. all interviews were conducted in bulgarian and later translated, although certain expressions in the original conversations were spoken in english. wherever that is the case, i have indicated so in the transcription. the group of interviewees includes people who were born in bulgaria between 1964 and 1974. all of them immigrated to the us after 1989 and all but one came as students pursuing 3 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 4 advanced degrees. three had completed doctoral degrees in the us and three others were enrolled in doctoral programs at the time of the study. the high level of education and the enterprising spirit of the people in the group no doubt had an influence on the life experiences they have had and the way they talked about them. in this sense, this group is not representative of all bulgarians and not even of all bulgarian immigrants in the us. however, all participants in the study spent their childhood, adolescence, and in some cases part of their early adulthood in bulgaria during the last years of the communist regime. this was a period of stability, characterized by a relatively egalitarian social organization for the majority of bulgarians. for example, education was free and widely accessible. high school education was mandatory. centralized structures permeated all aspects of social life and thus all young people had to go through certain collective experiences, such as participation in agricultural brigades, membership in communist youth organizations, or mandatory military service for all healthy men over the age of 18. it is not surprising that some of these common experiences emerged repeatedly in the narratives and served to bring the group closer together, while distinguishing it from “americans.” in this sense, the narratives of such experiences represent what can be considered “typical” experiences in the lives of many bulgarians growing up in the 1970’s. it is important to stress my positionality as a researcher in the process of collecting and analyzing the data. i am a member of the same generation of bulgarians as my informants and share some similar experiences to the ones they recalled. thus, my personal memories were a basis of comparison in analyzing their narratives and served as a measure for the authenticity of their stories. this type of insight may not have been available to an analyst of a different national and experiential background. at the same time, my identification with this generation and my own immigrant status in the us implicate my perspective as partial and one that carries an insider’s bias. however, my theoretical grounding is derived from a largely western tradition that i have come to know through my life and education in the us for the last eight years. thus, i attempt to maintain an analytical distance in the discussion of the data in order to give my conclusions significance that goes beyond an insider’s view. in this sense, i see this project as a bridging effort where i, as the analyst, adopt the role of a cultural interpreter seeking to make everyday life in communist bulgaria knowable outside of its local context. this project is only the beginning of a larger exploration into the nature of collective memories of communism and makes no conclusive claims. rather, it seeks to document and demystify to a broader audience the profoundly human experience of everyday life in a communist country. 4. discussion: demystifying life under communism in the next part of the paper i examine in turn the establishment of relevant affective communities and the construction of space and time in the memory 4 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/2 doi: https://doi.org/10.25810/tjqa-yy74 memories of everyday life in communist bulgaria 5 narratives of informants in relation to the identity work they accomplish through the acts of telling. 4.1. identifying affective communities the notion of community is often invoked in discussions of culture, memory, and collective identity, but pinning the concept down is a difficult task. in his theory of collective memory, halbwachs (1980) conceptualizes affective communities as groups of shared experience – for example, men who have fought together in a war, people engaged in a creative project together, people bound by familial relations, etc. he distinguishes between abstract communities, such as nations, which he terms “distant frameworks,” and groups of a more immediate nature, which he calls “nearby milieus” (76). the notion of distant frameworks is similar to anderson’s argument that national communities are, in fact, “imagined communities” which do not rely on direct interaction among their members (anderson 1983). by contrast, the idea of nearby milieus can be related to the theory of discourse through the concept of participation frameworks (goffman 1974, tannen & wallat 1987). halbwachs acknowledges that, “between individual and nation lie many other, more restricted groups. each of these has its own memory” (1980: 77). the fact that individuals participate in multiple groups and may occupy various roles within them speaks to the multifaceted and unstable nature of collective memories, which are shaped in each telling by the particular participation framework within which narrators are situated (cf. goffman 1954). communities of both “imagined” and “actual” types were referenced in the narratives of my informants. the first type referred to national identity (being bulgarian), a common culture, and a common language. in examining the narratives, i observed that belonging to a national community or a national culture was referenced most often in relation to symbolic artifacts, such as films, books, music albums, or other cultural texts associated with bulgaria. for example, the group often talked about and exchanged dvds, cds, and tapes of bulgarian films and music. several satirical comedies, produced during the communist period and starring bulgarian actor todor kolev, were among everyone’s favorites. the commonality among those films is that they depict everyday life in communist bulgaria without direct references to the political and ideological regime, yet poke fun at its absurdities. on more than one occasion, several informants remarked how difficult it would be to translate the particular humor that defines these films to foreign audiences. an often-repeated remark was, “how do you explain this to an american?” this comment suggests a tacit recognition among informants of a community of memory and identity at the broad level of nation and culture. the second type of communities referenced in the narratives included smaller groups of friends, colleagues, and relatives who share direct experiences. such references emerged when informants reflected on the meaning of direct personal experiences. consider for example the following excerpt from an 5 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 6 interview with val in which i asked him to talk about his memories of life in bulgaria before 1989. (1) example: communities of memory interviewer: 1 when you see your parents do you talk about how it used to be and things like that at all? val: 2 well in the interest of truth i talk about how it used to be with my friends more often 3 who are all my age 4 for example suddenly we would remember something which… 5 we weren’t appreciating at the time 6 and now suddenly when you go back it looks so typical although my question (line 1) asks whether val talks to his parents about the past, his answer identifies a different community of memory that is relevant to him, defined as “my friends… who are all my age” (lines 2-3). this disclaimer sets the stage for the narrative to follow and suggests that val constructs his memory of how bulgarian life “used to be” in reference to his generation and in particular to his group of friends. halbwachs recognizes the importance of generation for the maintenance of memory. he also notes that as communities change, so do their collective memories. when we lose our connection with a community of memory, we lose the memories associated with it. thus, for immigrants who find themselves separated from many groups they have left in the home country, it is important to create new affective communities with fellow expatriates within which stories about home and the past can be told and preserved. however, the types of stories that can be shared are limited to what is assumed to be “typical” or “common” among the members of the group, rather than deeply personal and private experiences. in this sense, the tellers rely on knowledge schemas (tannen and wallat 1987) to make judgments about the boundaries of the affective communities they form as immigrants. 4.2. narrative space: narrative constructions of “home” next, i examine the construction of space in the immigrant narratives. in his theorization of collective memory, halbwachs distinguishes between several types of space, among which are physical space, economic space, legal space, and religious space. for the purposes of this analysis, i define space simply as narrative references to a physical and/or symbolic location or place that situates the telling of the story. labov and waletzky (1967) have suggested that establishing place is one of the essential elements of any narrative. they propose that most narratives make use of five components, which they term orientation, 6 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/2 doi: https://doi.org/10.25810/tjqa-yy74 memories of everyday life in communist bulgaria 7 complication, evaluation, resolution, and coda. the orientation is typically found in the beginning of narratives and serves to “orient the listener in respect to person, place, time, and behavioral situation” (labov and waletzky 1967: 32, emphasis in original). my analysis of the memory narratives of immigrants confirms the importance of labov and waletzky’s orientation, although the examples discussed below illustrate that the construction of place is not always restricted to the beginning of narratives but can be interspersed throughout the story. in the immigrant narratives described here, the memory of life under communism is intertwined with the memory of home. “home” is a broad category that signifies both a physical location and a symbolic grounding, both of which are important to the teller’s identity. “home” is sometimes identified simply as bulgaria, the immigrants’ country of origin. at other times it is linked to particular locales, such as a teller’s hometown or place of residence. an interesting paradox arises, however, in the establishment of “home” as different from the narrator’s current location – “home” is designated as a distant “there,” different from the proximate “here.” “home” is an essential component in grounding the identity narratives because it provides a point of origin for the life stories of the narrators. in this sense, forgetting “where one came from” is a threat to the integrity of one’s identity and is valued negatively. an example of this can be seen in the excerpt below, where tina relates a story about a trip to chicago. tina took the trip with a bulgarian friend at a time when both of them lived in nashville, tn and neither had traveled to a larger city in the us. in the example below, tina recalls an exchange with her friend towards the end of their trip: (2) example: constructing notions of “home” tina: 1 so she said, “how i wish i didn’t have to go back to that village nashville.” 2 and i said, “village? you are from vakarel,* girl!” (laughs) 3 i mean, compared to vakarel, nashville is a much bigger town. steve: 4 where the fuck is vakarel? (everyone laughs) ((this sentence spoken in english)) *((vakarel is a small town in bulgaria, not distinct in any particular way.)) tina’s story, told in a group setting, provoked laughter and sarcastic comments from the people present. the brief interchange with steve has a strong evaluative component both in relation to tina’s friend, but also in relation to the meaning of “home,” especially as suggested by steve’s comment in line 4. both tina and steve implicitly position “home” as an obscure, rural place (lines 2 and 7 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 8 4) and at the same time poke fun at tina’s friend who seems to have “forgotten where she came from.” thus, the interchange establishes tina’s identity as someone who is better connected to home than her fellow traveler. at the same time, the story positions “home” as a place that is less sophisticated than “here” and reinforces tina’s present identity by implicitly justifying her choice to leave bulgaria. an excerpt from an individual interview with val also demonstrates how the narrative space of “home” is established in reference to a present location “away from home.” the excerpt contains a short narrative with the theme of remembering things that seemed “typical of bulgarian life as it used to be.” in this narrative, what was “typical” is represented in memory by everyday, consumer products. generally, the notion of “typicalness” was often indexed in the narratives through references to concrete everyday objects and products. on the other hand, the meaning of “the way life used to be” is clearly situated in spatial terms within the opposition of “there” (in bulgaria) and “here” (in the us). (3) example: spatial and material dimensions of “home” val: 1 i have explained this here 2 we have always laughed a lot 3 for example if they told you for example to buy vero* in bulgaria 4 i mean here if you tell someone 5 “go and buy palmolive or some other detergent” 6 you know, they know exactly 7 i mean in bulgaria vero was understood 8 as the only kind that is 9 detergent for washing dishes, simply there wasn’t another one 10 while here if you were simply told 11 “go and buy detergent for washing dishes”… 12 you will be in great difficulty *((vero is a bulgarian brand of dishwashing liquid. because of the lack of alternative products, the word “vero” was used by people instead of “dishwashing liquid.”)) this narrative talks about a very mundane experience – buying dishwashing liquid – but establishes this most trivial activity in terms of differences between “here” and “there.” the implied distinction is not simply geographical. the story describes a difference in the way of life in a consumer society “here” (in the us), and life in a society where consumer choices were much more limited (in bulgaria). the same narrative is visually represented below in a different way to demonstrate two parallel sub-narratives that show more clearly the distinction between “here” and “there”: 8 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/2 doi: https://doi.org/10.25810/tjqa-yy74 memories of everyday life in communist bulgaria 9 (4) example: spatial dimensions of home: “here” vs. “there” “there” (in bulgaria) “here” (in the us) 3 for example if they told you to buy vero in bulgaria 4 i mean here if you tell someone, 5 “go and buy palmolive or some other detergent” 6 you know, they know exactly 7 i mean in bulgaria vero was understood 8 as the only kind, that is 9 detergent for washing dishes, simply there wasn't another one 10 while here if you were simply told 11 “go and buy detergent for washing dishes”… 12 you will be in great difficulty these two sub-narratives illustrate the connection of memory to space, broadly defined. it is significant that the idea of memory space is often constructed in contrast to a physical space within which the narrator is located in the present. in that sense, bulgaria in the memories of the immigrants is a different bulgaria from the one that actually exists in the present moment. in the words of l. p. hartley (1953), “the past is a foreign country; they do things differently there.” 4.3. narrative time – “the way it was” space is closely related to the notion of narrative time. halbwachs outlines several different conceptualizations of time related to the personal experience of temporality and the more abstract notion of time as historical flow. time is particularly important to this study because “communism” can be thought of as a particular historical period. however, because the reconstructions of this period are accomplished through the means of narrative, ricoeur’s theory of narrative time is particularly useful. of specific interest is to this discussion is the distinction between what ricoeur terms the episodic and configurational dimensions of time. the episodic dimension structures narratives as a linear progression and characterizes a story as made out of distinct, sequential events. by contrast, the configurational dimension 9 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 10 of time implies a whole within which the sequence of events is significant or meaningful (ricoeur 1980: 178). the narrative acts of recollecting the “communist past,” then, are set within a configurational boundary of meaning. in outlining this boundary, the configurational dimension of narrative time also has an evaluative function that allows the teller and listener to make judgments about the episodic elements of the narrative. as an illustration of how a narrator may outline the configurational dimension, consider the following example from the interview with val. at the very beginning of our conversation val makes the following disclaimer without any prompt or question on my part: (5) example: configurational time and meaning val: 1 when you asked me, 2 when you requested this interview 3 i sat down and thought of several things that made an impression, 4 well, that i remember. 5 for example the way it was 6 at the time that we had to apply to the komsomol.* *((komsomol is a russian coinage that was appropriated in the bulgarian language to refer to the political organization of high-school and university students, which was an affiliate of the communist party.)) in my request for an interview, which i had made approximately a week earlier, i had told val i was interested in what people remembered about life under communism, and that i would like to talk to him about his own memories. i had hoped that by phrasing my interest in broad terms i would not lead my respondent in any specific direction. this strategy prompted val to set a particular configurational boundary around the past that made his own recollections meaningful. this boundary is introduced in lines 5 and 6. the phrase “the way it was” in line 5 also works as a narrative abstract in labovian terms (labov, 1972), or a summary of what is to come later in the story, and implies that the narrative to follow is authentic. in other words, val is making the claim that he is not recollecting a rare or unusual event but one that somehow typifies life under communism. several lines later in the narrative, val provides additional information about the period he has introduced in lines 5 and 6 above. in lines 10 to 13 below val makes assertions about the moral climate of the period and his statements exemplify the evaluative function of configurational narrative time: 10 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/2 doi: https://doi.org/10.25810/tjqa-yy74 memories of everyday life in communist bulgaria 11 (6) example: configurational time and moral judgments val: 10 i think absolutely nobody then already believed in anything 11 especially from our generation. 12 i mean, even our parents 13 i think had already stopped believing. this passage refers to people’s belief in the communist ideology, as defined by the ruling communist party, which was presumed to underlie individuals’ public, if not necessarily their private, behavior. by negating this presumption, val attempts to establish the past as “non-ideological” and, perhaps, more “ordinary” than an outsider would expect. however, it is interesting to note the use of the word “already” in line 10, which seems to imply an earlier configurational boundary – a time when “everybody believed.” while universal belief in communist ideology is not likely to have existed during any period in bulgarian history, the significance here is that val is working to construct through his narrative a period of “normality” within his memory of the communist past – a period that is free from ideological pressure by virtue of people losing their belief in ideological doctrines. a different illustration of how configruational and episodic time structure memory narratives is found in recollections of student life that are common to some degree among all informants. tina and leo, who are husband and wife and have known each other since high school, tell several such stories about high school life. their stories show a parallelism that comes from shared memories. the following opening lines introduce two narratives told in a group setting: (7) example: confugurational time and “typicalness” tina: i remember how we had to do group physical exercise every morning. (8) example: confugurational time and “typicalness” leo: i’ve had my hair measured. they used to measure your hair. these opening statements establish the theme, or the pattern of repetition that confers historical significance to the narrated experiences that follow. the telling of the stories themselves is accomplished through the episodic dimension of time. tina tells about mandatory physical exercise at school as a nuisance. leo tells a story about having to maintain a hair of a certain length in order not to be harassed by school officials. leo’s story reflects an aspect of life under communism according to which a teenager’s behavior was subjected to invasive 11 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 12 discipline even in the most mundane of circumstances. the story is an example of such disciplinary rules, as represented by the practice of “measuring your hair” at the entrance of school to determine if it met government issued decency standards. both tina and leo tell their stories in a humorous way and tend to exaggerate the events for comical effect. one result of this is that the “normal” past in their narratives is recast as “absurd” from the point of view of the present. consider for example another brief story told by leo in the course of the same group conversation. in this narrative leo recalls having to wear a high school uniform and the problems he had with the metal buttons on his suit: (9) example: configurational and episodic dimensions of time leo: 1 we had to wear these suits with metal buttons 2 and they always used to fall off 3 because the metal would just cut through the thread. 4 i mean it was idiotic that they made the suits with metal buttons 5 but you couldn’t replace them with other buttons. 6 so one day i got sick of sewing them back on all the time 7 and i used safety pins on the inside of the jacket to pin all of my buttons on. (everyone laughs) 8 i thought i was so smart. in this narrative leo establishes the configurational temporal boundary in line 1 and proceeds to tell the story along the episodic dimension in lines 2 through 7. line 8 provides a summative evaluation and establishes leo’s identity as someone who found a way of resisting the disciplinary rules. in this case, leo’s resistance can be interpreted as an attempt to reestablish “normality” within a context of absurd clothing rules. the laughter in response to leo’s story stems from the fact that everyone present at the telling can recall various problems with school discipline and uniforms. indeed, leo’s narrative prompts other people in the group to tell similar stories, which remained embedded in the same configurational dimension of time. in sum, the narratives i examined demonstrate that the configurational boundary of time is often established at the outset of the story. this technique implies that what follows is only one among a number of similar stories that are typical of life during that period and serves to evoke historicality. 5. conclusions: memory and identity in narrative in this essay i have attempted to document how a group of immigrants in the us remember everyday life in communist bulgaria. my study began with a 12 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/2 doi: https://doi.org/10.25810/tjqa-yy74 memories of everyday life in communist bulgaria 13 specific interest in recovering the memory of mundane experiences of “life under communism.” however, in the course of the study it became evident that memories of communism are difficult to isolate as distinct narratives. because communism was not simply a set of ideological directives but permeated nearly all spheres of social life, the theme of “life under communism” was intertwined in the memories of informants with other themes, such as “home,” “youth,” and “high school life,” to name a few. this layering of meaning is reflective of what goffman (1974) terms “laminations” and speaks to the complex and unstable nature of collective memory. in my analysis i have attempted to map the memory of communism emerging from the narratives onto the matrix of affective community, space, and time that was first identified by halbwachs in relation to collective memory. this process of mapping reveals three general trends in the way the narratives were constructed. first, the informants’ narratives of the past are acutely shaped by their present circumstances as immigrants. in this sense, the memory narratives contribute to a process of secondary socialization (berger & luckmann 1967) into the social universe of the us that each of my informants is undergoing. because immigrants are faced with the task of negotiating their identities against this unfamiliar social universe, their recollections of the past are often constructed in response to what they perceive as the nature of their immediate environment. thus, remembering life in communist bulgaria is often narratively accomplished by way of comparison to their present life in the us. at the same time, memory narratives often provide a narrative space of escape from the challenges of immigrants’ present lives and a way to reconnect with a distant homeland. second, the tellability (ochs & capps 2001) of immigrant memory narratives is linked to the narrators’ identity projects and the affective communities that contextualize the telling. for example, when leo recalls an act of small, personal resistance against the disciplinary rules of high school and presents the episode as humorous (example 9), he seeks to establish his personal identity as a free-thinking, independent human being. by contrast, when val describes the past as a period during which no one believed in communist ideology (example 6), he precludes the need for personal resistance because he constructs the social system as previously purged of ideological content. these differences illustrate the intersection of memory and identity as it occurs in personal narratives. thus, the meaning of “normal behavior” or “normal life” under communism is understood differently in the narratives of leo and val, and these differences are the result of the personal identity project of each narrator. finally, the meaning of “normality” in the past is also affected by the immigrants’ need to adapt to present circumstances. in their efforts to normalize the present and make it easier to cope with, narrators sometimes delegitimize the past and make it appear absurd or incomprehensible (see examples 4 and 9). this delegitimation is accomplished through an implied conversation with an imagined western audience, which is often simply referred to as “the americans.” yet, because of this implied conversation, immigrant narratives may be particularly 13 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 14 poignant as a way of representing and interpreting the experience of “everyday life under communism” to western audiences. because the construction of memory in narrative is a pragmatic process embedded in the context of the present and linked to various identity projects, the conclusions in this paper should not be taken as definitive or universal. the challenge for researchers interested in collective memory is to embrace the partial and unstable nature of their subject matter. unlike the work of historians who study archival materials, the goal of memory researchers is not to arrive at a fixed record of the past but, rather, to open the past up to multiple interpretations and voices. in this task, narrative analysis is particularly useful because it allows for the “systematic study of personal experience and meaning: how events have been constructed by active subjects” (riessman 1993: 70). by collecting the memory narratives of various groups, memory researchers can contribute to the recovery of a more democratic, multivocal version of the past. in relation to communism, a social system that is at once among the most maligned and most admired yet to exist, the project of recovering a fuller story of the past is all the more important. references anderson, b. 1983. imagined communities: reflections on the origin and spread of nationalism. london: verso. berger, p. l. & t. luckmann. 1967. the social construction of reality: a treatise in the sociology of knowledge. garden city, ny: anchor books. bruner, j. 1991. acts of meaning. cambridge, ma: harvard university press. burawoy, m. & k. verdery (eds.) 1999. uncertain transition: ethnographies of change in the postsocialist world. lanham, md: rowman & littlefiled. carey, j. w. 1989. communication as culture: essays on media and society. boston: unwin hyman. gagnon, j. h. 1992. “the self, its voices, and their discord.” in c. ellis and m. flaherty (eds.) investigating subjectivity. newbury park, ca: sage. goffman, e. 1959. the presentation of self in everyday life. garden city, ny: anchor books. goffman, e. 1974. frame analysis: an essay on the organization of experience. cambridge, ma: harvard university press. halbwachs, m. 1980. the collective memory. new york: harper & row. hartley, l. p. 1953. the go-between. london: h. hamilton. holstein, j. a. & gubrium, j. f. 2000. the self we live by: narrative identity in a postmodern world. new york: oxford university press. labov, w. 1972. “the transformation of experience in narrative syntax.” in w. labov (ed.) language in the inner city: studies in the black english vernacular. philadelphia: university of pennsylvania press. labov, w. & j. waletzky. 1967. “narrative analysis: oral versions of personal experience.” in j. helm (ed.), essays on the verbal and visual arts. seattle: university of washington press. pp. 12-44. 14 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/2 doi: https://doi.org/10.25810/tjqa-yy74 memories of everyday life in communist bulgaria 15 ochs, e. & l. capps. 2001. living narrative: creating lives in everyday storytelling. cambridge, ma: harvard university press. ratushinskaya, i. 1988. grey is the color of hope. new york: knopf. ricoeur, p. 1980. “narrative time.” critical inquiry 7(1). pp. 160-190. riessman, c. 1993. narrative analysis. newbury park, ca: sage. sherbakova, i. 1992. “the gulag in memory.” in luisa passerini (ed.) memory and totalitarianism. international yearbook of oral history and life stories. vol. 1. oxford: oxford university press. pp. 103-115. tannen d. & c. wallat. 1987. “interactive frames and knowledge schemas in interaction: examples from a medical examination/interview.” social psychology quarterly, 50(2). pp. 205-216. todorov, t. 1999. voices from the gulag: life and death in communist bulgaria. university park, pa: the pennsylvania state university press. verdery, k. 1996. what was socialism, and what comes next? princeton, nj: princeton university press. williams, r. 1977. marxism and literature. oxford: oxford university press. 15 kaneva: memories of everyday life in communist bulgaria published by cu scholar, 2006 colorado research in linguistics 6-2006 memories of everyday life in communist bulgaria: negotiating identity in immigrant narratives nadia kaneva recommended citation microsoft word !kaneva_cril_2006_final.doc the cognitive correlates of aspect and the english present perfect 1 raymond: the cognitive correlates of aspect and the english present perfect published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/7 doi: https://doi.org/10.25810/v7m1-5p21 3 raymond: the cognitive correlates of aspect and the english present perfect published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/7 doi: https://doi.org/10.25810/v7m1-5p21 5 raymond: the cognitive correlates of aspect and the english present perfect published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/7 doi: https://doi.org/10.25810/v7m1-5p21 7 raymond: the cognitive correlates of aspect and the english present perfect published by cu scholar, 1995 8 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/7 doi: https://doi.org/10.25810/v7m1-5p21 colorado research in linguistics 1995 the cognitive correlates of aspect and the english present perfect william raymond recommended citation tmp.1537648250.pdf.d1nxi gossiping: introducing direct quotes in story retelling 1 taimi metzler: gossiping: introducing direct quotes in story retelling published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/6 doi: https://doi.org/10.25810/8b4s-da74 3 taimi metzler: gossiping: introducing direct quotes in story retelling published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/6 doi: https://doi.org/10.25810/8b4s-da74 5 taimi metzler: gossiping: introducing direct quotes in story retelling published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/6 doi: https://doi.org/10.25810/8b4s-da74 colorado research in linguistics 1995 gossiping: introducing direct quotes in story retelling s. taimi metzler recommended citation tmp.1537649198.pdf.wdqx9 education reform and language politics in the coroico municipality of the nor yungas of bolivia education reform and language politics in the coroico municipality of the nor yungas of bolivia∗ victoria stockton university of colorado aymara, one of four national national languages in bolivia, has become endangered within the past generation in the coroico municipality. top-down education reforms implemented in 1994 have adopted a language-as-resource orientation to alleviate the degradation of bolivia’s indigenous languages. bottom-up grassroots movements nationwide reveal a tenuous shift away from colonial-era language attitudes. the gap between the language policy of bolivia, as enacted by the education reform, and the practice of that policy at the grassroots level characterizes the contentious and shifting social atmosphere of bolivian sociolinguistic culture. my focus centers on historical legislation and language attitudes against multilingualism, as well as legislation and language attitudes promoting multilingualism. this case study exemplifies efforts to curb language abandonment in the face of globalization and the growth of world languages. 1. introduction the coroico municipality, approximately 60 miles over the andes from the capital city of la paz, bolivia, encompasses a majority population of aymara semi-subsistence agriculturalists. one of bolivia’s largest indigenous groups, the aymara in the coroico municipality cobble together an existence at once remote and global. the town of coroico, with 3,500 inhabitants, is an attractive and scenic tourist destination, the favorite of many international tourists and residents of la paz on weekend holiday. a modern highway, completed in the last few years, takes passengers (more) safely over the andes from la paz to coroico. but the steep, semi-tropical hillsides of the nor yungas, the region of which coroico is a part, keep other towns within the municipality isolated. in essence, the geopolitical makeup of the coroico municipality reveals its sociocultural complexity: the coming-together of multiple ethnicities, subsistence patterns, socioeconomic standings, and languages. the two most commonly spoken languages in the nor yungas are aymara and spanish, the latter being the language of the government and state. colonialism historically attempted to eradicate indigenous languages to unify the ∗ my deepest gratitude to carol conzelman, whose guidance, knowledge and support made my research in bolivia possible. thanks to kira hall and adam hodges for their support, ideas and insight. colorado research in linguistics. june 2005. vol. 18, issue 1. boulder: university of colorado. © 2005 by victoria stockton. 1 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) country and its indigenous people under one language: spanish. government legislation during this era enacted education policies to punish public use of indigenous languages, including aymara. currently, with bolivia’s neoliberal reforms—which knit the nation more closely to the global market—and impending increased tourist traffic to the region as a result of the new highway, knowledge and use of spanish has become increasingly important as the aymara agriculturalists in coroico utilize and participate in the democratic reforms of their country. by the same token, government legislation resulting in the law of popular participation, which includes the education reform of 1994, enacts a promotive language policy requiring the presence of bilingual education in regional schools where more than one language is spoken by the community. the education reform reveals efforts by the government to elevate human rights and prevent the loss of linguistic capital in bolivia. aymara, therefore, maintains a tenuous existence as a language. simultaneously dying out and being revitalized, shunned and used with pride, the contextual contradictions of aymara use and language attitudes encapsulate changes within the larger social milieu of the country. this article will discuss the interaction and effects of education reform, globalization, and minority identity politics for sustainable multilingualism in the coroico municipality of bolivia. the discussion will take place in two parts: the first part will discuss the preponderance of language shift in the coroico municipality, which may lead to the death of aymara in the municipality within two generations. to understand the phenomena of language shift, an understanding of historical social factors motivating negative language attitudes1 is required. i will then discuss the role of globalization and bolivia’s adoption of neoliberal reforms and how they influence language politics in the country. the second part of the paper will discuss the potential reversal of language shift in the coroico municipality due to promotive education policies and social movements that valorize indigenous languages and a multilingual nation. the education reform of 1994, as part of bolivia’s transition to a capitalist democracy, is a policy that signifies remarkable transformation for the national economic stability of bolivia, the vitality of indigenous languages, and indigenous rights. as a pivotal turning point to reverse negative language attitudes and the loss of aymara, the reform attempts to promote indigenous languages and cultural diversity, which would alleviate bolivia’s endemic poverty by creating more equitable opportunities for citizens. in order to be successful, the reform must replace negative language attitudes and social stigmas that view 1 i use the term “negative language attitudes” to refer to individuals or communities that feel a language marks them as lesser members of society, thereby wishing to discontinue the use of the language in public spheres. 2 2 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics language as a problem with the attitude that multilingualism is a social advantage, viewing language as a resource. if the community overcomes the negative social stigmas associated with their language, social mobility will not be hindered by the public use of aymara. both the efficacy of the reform and the ways in which people incorporate it into their lives depend heavily on sustained implementation and popular participation. the language attitudes within the community are divided between those that feel aymara is valuable and integral to bolivian culture, and those that feel aymara is useless to citizens negotiating their role in modern society. the reform will not be effective if the polarity of language attitudes in the municipality is not reconciled. i analyze how language use, practice, policy, and social stigmas expose the gap between the bolivian government and the yungueños in their desires for language survival, modernity, and effective democracy. the coroico municipality exists in a precarious sociohistorical moment. the community faces the extinction of aymara due to language shift and negative attitudes associated with aymara use. but as the effects of the education reform catches on, and as aymara and other indigenous languages are used in public more frequently, negative attitudes may recede and pave the way for multilingualism to become a sustainable reality. this paper offers a sociohistorical analysis of the terrain of language attitudes in the coroico municipality and the ideologies that influence those attitudes. i spent three months in bolivia researching the education reform and language politics in the coroico municipality. participant observation as well as structured and semi-structured interviews informed my findings. some interviews were scheduled ahead of time and tape-recorded. i also conducted semistructured interviews, but recorded them only with field notes due to formalities with my interviewee. i conducted informal interviews when i would participate in an event and take the opportunity to chat with people in that setting. unless i had permission to tape record ahead of time, i relied on copious field notes. i selected interviewees in a snowball-like fashion: one person telling me about another person, telling me about another person, and so on. i know that i was unable to interview several very important figures whose opinions and experiences would have greatly contributed to this article. i am also sure i inadvertently left out several others whose expertise or experiences i never knew or heard about. before and after the completion of my fieldwork, i spent three months doing background research with secondary sources to deepen my understanding of the situation i was entering into and my findings once i completed fieldwork. 2. language shift in coroico at the one-room school in the town of chacopata in the coroico municipality, the students were asked which of them could speak aymara. 3 3 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) nobody answered; silence gave way to hushed giggles concealing embarrassment and shyness. the children all pointed to one student, reporting in spanish that he spoke aymara. the student put his head down, shaking it, and denied the claims. in the coroico municipality the majority of parents speak aymara as their first language, especially in the rural areas. these parents actively choose not to teach their children aymara and discourage them from admitting their aymara heritage. parents speak the language only between each other, behind closed doors, refrain from speaking aymara in public, and inform local teachers that they want their children instructed only in spanish, despite knowing their neighbors speak both languages. the coroico municipality is undergoing a first-generation language shift. in the past sixteen to twenty years, parents stopped teaching their children aymara as the first language, replacing it with spanish. the non-transmission of the mother tongue to children bodes poorly for the survival of aymara in subsequent generations. the fact that this generation of children was raised without aymara indicates that they will most likely not raise their own children with aymara, and within two generations the language will have completely disappeared from the nor yungas. aymara parents have become embarrassed about teaching their children aymara in a region of predominantly aymara communities for different reasons. parents, school officials, teachers, and students expressed that first-generation language shift is occurring in coroico because aymara is only useful for speaking to one’s grandparents or parents, nowhere else in life. one student, a young man attending the rural university in the coroico municipality described to me how he was raised: en la casa, aprendemos el aymara de los papas. más que todo, por ejemplo, mis papas hablan el aymara entre si. pero cuando me hablan a mi, me hablan en castellano. y hablo en castellano para responderles. pero, hay mi abuela, por ejemplo, mi abuela sólo habla aymara. ella no puede hablar en castellano. entonces, ella me habla en aymara, y entonces yo también la contesto en aymara. así no comunicamos. es que, la lengua aymara nos facilita comunicarnos. o sea, sólo se entiende entre nosotros. y ¿qué tal si voy a otro país? el aymara no va a servir para nada. en eso piensan los padres también. no hay muchas personas que hablan en aymara. yo creo que, la lengua aymara es una lengua que nos facilita comunicar, más que todo, con las personas mayores. (jorge, personal interview, march 2004) we learn aymara from our parents. not in school. more than anything, my parents speak aymara between each other. but when they speak to me, they speak in spanish. and i speak in spanish when i respond to them. but my grandmother, for example, she does not understand spanish. so, when she speaks to me in aymara, i answer her in aymara. that is how we communicate. aymara helps us communicate between each other. that is, it’s how we understand each other. but what if i want to go to another country? aymara will not be useful at all. not many people speak aymara. i believe that 4 4 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics the aymara language is a language that facilitates communication, more than anything, with older people. (author’s translation) jorge describes how his parents would speak aymara only with each other yet actively chose to speak spanish (castellano) with him. he feels that speaking aymara is useful only for speaking to older people, and will not be useful in the future or for anything outside of communicating with elderly family members. as many analysts of language vitality have illustrated, if young speakers feel that a language is useful only for speaking with grandparents and have no desire to speak it amongst themselves, let alone to their own children, language death is imminent. in another interview, a school administrator, raul, described to me the class system in bolivia. he told me that there are three classes, the upper class, the middle class and the lower class. the indigenous people—aymara, quechua, guaraní—are all lower class. in order to ascend in society, the indians have to get rid of the social markers that mark them as lower class, backward, uneducated. so they deny their language. they teach their children spanish so that they can go to school and be successful; so they will not be labeled as indians based on the language that they speak. according to raul, social class is not conflated with ethnicity at any level other than lower class. bolivians want to assimilate to the upper classes systems, and in order to learn about computers and politics and medicine, they have to know spanish. aymara is no longer useful to the people because none of the businesses use aymara, and the people do not want the social stigmas attached to the language hindering them. these narratives reveal more than just the attitude that aymara is not useful for anything. obviously, a language will be transmitted from generation to generation only if it is useful to people in everyday life and in multiple arenas of life, not just for speaking with one’s grandparents. a language, above everything else, exists to facilitate communication between people. when that function no longer exists, the language need not exist. but nothing about the language, as a formal system of communication, makes it better or worse than other languages for allowing people to communicate. thus, it is the social values attached to the language that create the perception among speakers that the language is not useful. people in coroico often said that aymara is not used in public because of embarrassment or shame. using it marks them as backward, degenerate, uneducated: of a lower class. the embarrassment and shame attached to the use of aymara in the public sphere was described to me in detail by one man i interviewed, carlos, a teacher involved with literacy programs among adults in the coroico municipality: cuando van, por ejemplo, a la paz, y quieren visitar a un ministro, si hablas aymara, te sacan empujones. pero si hablas en castellano, te dicen ‘pasa, toma asiento’. sí. pero si hablas aymara, te dicen, ‘aaah, ¡afuera afuera afuera!’. entonces, yo creo que, esta experiencia amarga hace que la gente se asume de 5 5 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) esta manera. dicen que no vas a sacar más alla, aprendiendo el aymara. tienes que hablar en castellano, y sólo así vas a salir bien. y si sóla hablas aymara, hay discriminación, hay marginalización, que viene desde la clase dominante. y, vas a una oficina en el sector público, y hablas en aymara, te van a buscar un traductor. ‘¿quién sabe hablar el aymara?’ pero la persona que grita esta pregunta sabe hablar aymara, entiende aymara. pero, ¿sabiendo que entiende, no? pero, tiene verguenza de, tiene verguenza de sea identificado como aymara, como campesino. y, bueno, es puestamente producto de educación. de sistema educativo. el sistema educativo hace que tengamos verguenza de nuestra cultura, de nuestras costumbres, nuestras tradiciones. los campesinos han estado siempre marginados, ¿no? campesino es cualquiera que vive en el campo, ¿no? que no quede de encima, entonces, yo creo que esto se viene desde la época de colonialismo. la marginación. y ahora digamos la ignación, digamos, en que de tener verguenza, para mí es producto de este sociedad en que vivimos. más de, el sistema educativo. porque creemos que la educación reintero, no, es un sistema muy poderoso. (carlos, personal interview, march 2004) when they go, for example, to la paz, and want to visit a ministry or an office, if you speak aymara, they shove you. but if you speak in spanish, they tell you, ‘come in, take a seat!’ yes. but if you speak aymara, they say to you ‘aaahh, out out out!’ so, i believe that this experience, it creates bitterness. they say, then, that you aren’t going to get anything out of learning aymara. you have to speak spanish, and only then things go well for you. and if you only speak aymara, there is discrimination, marginalization, that comes from the dominant class. if you go to an office in the public sector, and you speak aymara, they tell you to find a translator. ‘who knows how to speak aymara’ they yell, but the person who just yelled the question knows aymara, can speak aymara. they know and understand, no? but they are embarrassed, embarrassed to be identified as aymara, as a peasant. it’s a product of the previous education system, the education system created in us an embarrassment of our culture, of our customs, our traditions…the peasants have always been marginalized, no? a peasant is someone who lives in the countryside, no? that’s what the word means, but now it has come to encompass all indigenous people, whether they live in the countryside or not. it means they can’t rise to the top. i believe this came from the colonial epoch. the marginalization. and i believe that the embarrassment stems from this, and is a product of the society in which we live. more, the education system, because we believe that education reiterates those values. it is a powerful system. (author’s translation) carlos describes the humiliation and rejection that people experience when they speak aymara in public, and the extent to which people will hide knowing the language. he then attributes the persistence of the embarrassment and marginalization to the education system of the colonial era, saying it had a powerful role in shaping social consciousness. the phenomena of language shift in the coroico municipality is the product of socio-political conditions in bolivia that have, through centuries of colonization, created hierarchies of identity, prestige, social class, and mobility. 6 6 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics 2.1 the legacy of colonialism the potential for language death in coroico exists in part due to the legacy of linguistic nationalism. historically, the bolivian government attempted to eradicate indigenous languages in favor of spanish to cohere the nation as a monolingual whole. schooling was thus conducted only in spanish, and use of any other language in public shamed the individual and marked them as uneducated and low-class. in bolivia, the language policy of spanish-only in the colonial era produced and reproduced the social stigmas attached to the use of indigenous languages in the public sphere (luykx, 1999; albó, 1999). speakers of indigenous languages remained low in the social hierarchy and could only legitimately access the public sphere through the use of spanish. the policy, and the ideology behind it, was applied to the school system for the purpose of shaping citizens to ascend to the dominant minority. schools operated as ‘civilizing’ institutions (luykx, 1999), with the intent of erasing cultural and linguistic difference among the indigenous populations. schools instructed students in spanish only, regardless of students’ mother tongue, and focused on rote learning and memorization. students were punished publicly for speaking in their mother tongue at any time during school hours. xavier albó, a bolivian linguist, writer, and anthropologist, has written several books discussing bolivia’s education reforms, language policies, and ethnic movements throughout the past thirty years. an anonymous source quoted in an albó (2003) text describes his experiences in school: yo tenía un profesor lammado c., que vive hasta ahora. cuando yo hablaba en mi idioma aimara me mandaba a la cancha y en las dos manos nos ponía piedras y nos hacía alzar un pie. un centinela vigilaba y me golpeaba con el palo cuando bajaba el pie, todo por hablar mi idioma. así yo viví. (narrative in albó, 2003:31) i once had a teacher named c., who is still alive today. when i would speak in my language, aymara, he sent me to the schoolyard and in my two hands he put rocks and told me to raise one foot. another student would stand watch and would hit me with a stick if i let that foot fall, just for speaking my language. this is how i grew up (author’s translation) as albó (2003) illustrates, anyone who spoke an indigenous language, i.e. aymara, quechua, or guaraní, were punished in school because those who spoke indigenous languages were considered backward, uneducated, ignorant, and lower class. those who spoke spanish, and spoke it without an indigenous accent, could not be labeled as such and were accepted into society. only through the use 7 7 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) of spanish was it possible to excel in school and in vocation, and move into the upper echelons of society. school curricula were then designed by officials who maintained that lower class citizens, in order to ascend to true bolivian status, needed to speak spanish and forget their ‘backward’ ways of life: in 1954 the international labor organization of the united nations began in bolivia its first action program on behalf of native peoples anywhere in the world…deputy director of the ilo, jeff rens [said] the objective of the program was indian integration “by making a single people of two populations separated by origin, language, and way of life…in their eyes an educated indian is no longer an indian, he has become a man. (healy, 2001 quoting rens, 1961) students, socialized into thinking their language and culture separated them from dominant society and from being considered ‘fully human,’ associated negative values with indigenous language use and felt that they were nothing if they could not speak spanish. the symbolic domination of covertly requiring speakers to access a certain mode of speaking in order to be socially acceptable is both produced and reproduced sub-consciously: produced by the dominant ideology interlaced within the education system and reproduced by those who feel they need to change their mode of speaking in order to access a certain domain of life (bourdieu, 1991). if indigenous aymara speakers did not learn spanish, they had no way to communicate with anyone in the ‘legitimate’ domains of society, much less actually enter those domains. 2.2 language and nationalism punishments in school, such as the one álbo (2003) quotes in the above passage, functioned to create a negative association in children’s minds about using their mother tongue in public. these methods were meant to unify the bolivian populace and cohere them as a nation in the aggressive promotion of spanish in the classroom. linguistic similarity acts as a cultural marker that puts tangible boundaries around an otherwise imaginary community, constructing and legitimizing the nation and providing a basis for nationalism (anderson, 1983). language has always been seen as fundamental to building and establishing nations, especially since the end of the eighteenth century. by default of this process, linguistic minorities result from the nationalism that bars them from full participation in the state (heller, 1999). language as a defining factor of community boundaries implies that one can be associated with or dissociated from the community by virtue of language use. the association is not projected by any inherent properties of the languages in conflict, but by communities using them for social ends. the process of using language as a tool for nation-building results in subjugation and discrimination of minority groups whose language differs from the national language. 8 8 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics the deterministic ideology within bolivia’s previous education system—a covert policy suggesting that the ideal citizen, as a bolivian, speaks only spanish—created a polarity of identity based on language use. as bourdieu (1991) suggests, language use is founded on social laws of construction, and the construction of an “official language” establishes a hierarchy of linguistic practices. any mode of speaking that deviated from the ‘standard’ or ‘official’ language was socially measured at a degree lower in value than that standard. a unified linguistic ‘market’ is essential to the formation and legitimacy of the state, and the education system functions to shape citizens as competent in the legitimate mode of expression. linguistic competence then becomes the rubric against which educational progress can be measured. thus, language use ultimately reflects social distinction rather than linguistic distinction. the negative language attitudes created in the colonial era remain embedded in the social fabric of bolivia today, as witnessed by the reasons community members, especially parents, offer for not transmitting the mother tongue to children. however, education methods and social stigmas from the colonial era do not explain why language shift has only just begun in coroico. current changes in the bolivian government, society, and the increased tourist traffic to coroico motivate the language shift prevalent in the past generation. 3. globalization and bolivia’s democratic reforms bolivia passed the law of popular participation and other neoliberal reforms in the 1990’s in an attempt to integrate with the global market economy. the internationalization of economic, industrial, and technological resources in bolivia has ushered in the era of globalization. for bolivia’s languages, globalization is a double-edged sword. on one side, it exacerbates the process and rate of language death as indigenous languages are subjugated not just on the national level, but also on the international level (phillipson, 1992; skutnabbkangas, 2003). world languages, like english, have broadened their international sphere, expanding the locus of their function and use. more and more englishspeaking and spanish-speaking tourists and businesses enter into the lives of bolivians in rural areas, like coroico. computers and the internet also make the global more intimate with the local. thus, globalization and bolivia’s neoliberal reforms provide a more urgent economic impetus for parents to teach their children spanish as their first language and abandon aymara. however, current globalization theorists who focus on the effects of globalization emphasize that other globalizing trends reveal opportunities for diversity and expression (giddens, 1990; appadurai, 1996; hornberger, 1998; heller, 1999; brysk, 2000). in bolivia, neoliberal reforms and the transition to a global market economy were designed, in part, to alleviate the endemic poverty of the nation. reforming the education system and human rights policies for the minority populace have proffered the space for indigenous social movements and linguistic revitalization. indigenous language groups cultivate self-determinism 9 9 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) in part because globalization “creates new institutional links across borders, such as international organizations, integrated markets, and transnational social movement networks…globalization privileges the role of information and communication…all of these changes grant new access to power, as they voice identities and messages across borders” (brysk, 2000: 11). the possibility for people to network with others across borders and nations without geographical constraints redefines the territory in which power relations operate. although globalization currently motivates the first-generation language shift in coroico, it also, through neoliberal reforms, has the power to prevent the loss of aymara. linguistic revitalization movements and the 1994 education reform work to subvert the social stigma and negative attitudes of indigenous languages in bolivia, offering a potential reversal of language shift in the coroico municipality. 4. changes in bolivian language ideology during the celebration of bolivia’s día del mar, the day of the sea, school children in coroico stood in formation with flags, posters, and props to rally for the national goal of reclaiming bolivia’s lost sea coast. throughout my time in coroico, i had only heard aymara used in public by adults, and usually during political meetings in the countryside. however, on the day of the sea i was taken by surprise when the multi-colored flag representing the indigenous nation of bolivia came to the fore and children, looking immaculate in their school uniforms, presented speeches and songs in aymara. it surprised me because i had learned, by that time, that public use of aymara, especially among children, was shunned. surely the aymara message was lost on the monolingual spanishspeaking students and townsfolk, which made the display of aymara during this public celebration notable and symbolic. the use of aymara in the day of the sea celebration illustrates the complexity and shifting atmosphere of language politics in coroico. indigenous social movements in bolivia have laid claim to language as a vital cultural marker of the indigenous group, and the social demands made by these movements often entail abolishing the negative social stigmas attached to language use. following ethnic revitalization movements in the 1950’s, 1970’s, and the emergence of democracy in the 1980’s, new language policies began to emerge that respected the multilinguistic reality of bolivia. the government began incorporating multilingualism and respect for indigenous languages into its nationalist ideology. 4.1 language policy and the education reform of 1994 in 1994, gonzalez sanchez de lozada, in his first term as president, implemented the plan de todos, including the law of popular participation and the education reform, which proposed a restructuring of school systems, 10 10 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics materials, and teacher training to incorporate bilingual and intercultural education. the law of popular participation redistributed the political and administrative boundaries in the country and increased budget allocation to municipal governments to 20%. as 85% of the municipalities are comprised of indigenous majorities, the redistribution of boundaries and funds potentially empowers the rural indigenous peasantry to choose how their resources are allocated and how to accomplish governance on their own terms (healy, 2001). previously, municipalities were governed by centrally-appointed officials who traveled from the capital and instituted the rule of law. the lpp made decentralization more democratic and community based, instituting participatory planning and incorporating indigenous cultural practices and vigilance councils (healy, 2001), allowing citizens in each municipality the opportunity to participate in determining the form and quality of their government. in 1999 the bolivian government granted official national language status to aymara, quechua, and guarani alongside spanish. the designation of the three most widely spoken indigenous languages as national languages can be seen as both a response to the growing demand for indigenous legitimacy and cultural pluralism, and an initiation to reverse the trend of language loss. the 1994 education reform with its bilingual education component “aims to halt the decline in indigenous language fluency in the younger generations and raise aymara, quechua, and guarani to the status of truly ‘official’ languages” (luykx, 1999: 13). language policies act as political tools for shaping and maintaining the polity they represent, which emerge from the linguistic culture in which they function. by qualifying aymara, quechua and guarani as national languages, macro-level governmental decree creates an overt multilinguistic reality. because language is often a cultural marker by which nationalisms are justified, language policy either reflects the sociocultural reality it is grounded in or attempts to create it. analysts of world language policies examine the ‘fit’ between the policy and the polity it operates within (schiffman, 1996). for any language policy to be qualitatively understood, it cannot be divorced from the group of people it governs. language policies do not emerge a priori, but are constructed to produce and reproduce the various ‘rules’ by which speakers engage to access legitimate domains of speech. types of speech, knowing when and how and where to speak in certain ways and to certain people are covertly understood by members of a given polity, and the social values attached to various speech codes often arise from the language policy that officiates the public sphere. speakers acknowledge the overt statements of the language policy and its underlying values to negotiate the formation of identities as citizens within that polity. for example, the covert monolingual policy of the united states, establishing english as the dominant language but never explicitly saying so in a legal document, creates a hierarchical formula. as the language that all citizens must access in order to function in places such as business or school, english occupies the top rung of the hierarchy and other languages fall below, such that 11 11 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) speakers of languages other than english, although they do not have to give up their language, must also learn english in order to weave into the social fabric of the united states. native english speakers, by fiat, have no obligation to learn any of the other languages that are spoken in the united states due to the implicitly understood and reinforced ‘law’ that english is the accepted standard by which citizens communicate. a language policy, in varying degrees, either ignores the multilingualism of the nation in order to create a monolingual state, or reflects the multiplicity of codes within the state, conceding diversity as the marker of the national polity. the way a policy fits with its polity reveals much about how language is used to shape a sociocultural reality, and the nature of the linguistic culture in which it is grounded. therefore, the language policy of bolivia reveals a great deal about the shifting social attitudes about indigenous language use in the public sphere. the current promotive bilingual policy of bolivia draws upon a languageas-resource orientation which regards multilingualism as an advantage rather than a disadvantage. the vision behind a language-as-resource orientation is one of pluralist pragmatism, in which language becomes capital to its users rather than an emblematic tool of exclusion and nation-building (anderson, 1983). the more linguistic codes one can access, the more power one has. in bolivia, and in coroico especially, this can be exemplified by aymara politicians using both their indigenous language to communicate and identify with other aymara speakers, while also allowing them to communicate and identify with wider national politics through their use of spanish. the ability to speak more than one language in a multilinguistic society puts users at an advantage by allowing them to access multiple social groups and to identify with more people. in this sense, language is capital and identities are negotiated via the application of that capital in differing social contexts. language is not inherently exclusionary, but it has been used to those ends, especially for the purposes of nation-building. shifting the discourse to a language-as-resource orientation, the exploitation of linguistic capital builds relations between groups, and the once-marginalized group escapes inevitable social subordination as the stigmas attached to indigenous language use deteriorate. the education reform of 1994 attempts, among other things, to reverse the internal colonialism achieved in previous education methods. the current social and economic conditions contribute to a context in which educational discourse shifts to one of pluralism, of unity succumbing to diversity, so that multilingualism is no longer an obstacle to national unity, but its descriptor. 4.2 reversing language shift in coroico due in part to the social changes in bolivia in the past thirty-five years, the public use of indigenous languages has amplified in the past decade. aymara can be heard on university campuses, in classrooms, and during political 12 12 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics demonstrations and national holidays, like the day of the sea. language attitudes are changing nationwide, but the efficacy of the education reform in the coroico municipality is yet to be determined. the reform must contend with existing negative language attitudes and the current first-generation language shift. reversing language shift and altering language attitudes are vital to creating sustainable multilingualism and curbing language death in the region, but without reception from the community, the education reform is a meaningless governmental decree. 4.2.1 the unidad academica campesina the initial step toward creating bilingual education requires the availability of trained teachers in the bilingual modality, as well as texts and materials for bilingual instruction. at the unidad academica campesina (uac), a local university affiliated with the bolivian catholic university in la paz, both three and five year programs are offered to train teachers in the bilingual education modality of the current reform. applied three years ago, the pedagogía program in the uac is the first program in the yungas to offer this training, and students from all over bolivia attend the university so that they can become bilingual teachers. the pedagogía program requires students to learn both the methodology of teaching bilingually and also to create bilingual texts. despite the process of language shift in the coroico municipality, each of the students in the pedagogía program felt that through the implementation of the reform in the local schools, the erosion of indigenous languages would abate: los profesors nos enseñan en dos lenguas, que son el aymara y el castellano. y este es el bilinguismo, que para nosotros es muy importante porque nosotros vayamos allá a los cultos, y plantamos a los niños a enseñarlos en las dos lenguas. y yo creo que es bueno llevar esto para nuestro futuro, aquí, para no perder nuestra cultura, y ir adelante, y mejorar la educación (jorge, personal interview, march 2004) the professors teach us in two languages, aymara and spanish. and this is bilingualism, which for us is very important because we will become more educated and we plant this in the children that we will teach in both languages. and i believe that it is good to have this for our future, here, so that we don’t lose our culture, and so we can go forward, and improve the education, and maintain the languages. (author’s translation) jorge extends the philosophy of bilingualism from the training teachers undergo to how they will conduct their classrooms. he thinks of the future, of the role bilingualism will play for the students and their education. more than just a tool for better comprehension, jorge describes how bilingualism will revitalize the culture, maintain the languages, and bring forward the indigenous populations. 13 13 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) another student, felipe, also described the benefits of bilingualism in terms of resolving the social stigmas behind the use of indigenous languages. si va a estar en todas las partes, ya aplicando, ya como, aquí como lo están enseñando sí podría resolver la discriminación y verguenza. porque van a saber de sí, digamos, a valorarse ellos mismos que son de origen aymara y también saben hablar castellano. si no lo vamos a enseñarles, creo que las lenguas originarias van a morir. en cambio, enseñandolos a los niños ellos van a mantenerlo, la lengua. (felipe, personal interview, march 2004) if [the reform could be implemented] in all parts of bolivia, all parts of it applied, like how it is being taught to us here, it could resolve the discrimination and shame behind use of the indigenous languages. because they are going to know how to valorize themselves, where they come from and they will also know how to speak spanish. if we don’t teach the children in their native tongue the indigenous languages will die out. if they don’t use their language, it will be lost. in change, teaching the language to the children will maintain it. (author’s translation) each of these students comes from different regions of bolivia with different linguistic backgrounds, one a native quechua speaker and one a native aymara speaker. they have plans to return to their hometowns and teach at the local schools. the five-year program requires that students write a thesis on original fieldwork, after spending time in the local communities and actively working with them. the students i interviewed acknowledged that, although the bilingual parts of the reform are not currently implemented in the coroico municipality, change comes step by step. when i asked jorge about the lack of support and resources for the local schools to have bilingual education, and how many teachers in town felt that it would not happen, he replied, “yes. but as it would be, we are in the process.” the teachers i spoke with also emphasized that bilingual and intercultural education will soon come to the yungas, but first the teachers must be trained and materials produced before any change can be seen. one teacher, professor luchaqui, stressed that the current teaching program had only arrived at the uac three years ago. this year they will graduate the first teachers licensed to teach bilingual and intercultural education, and subsequent years will see more and more teachers bringing the program to the local schools. he explained that, as part of decentralization and the law of popular participation, each district in the municipality has a director who oversees the programs and the methods of teaching in each school. the director analyzes the linguistic makeup and needs of the communities, and works with the teachers and the parents to formulate a curriculum that will fit the needs of the children. the teaching program at the uac forecasts the lasting success of the reform, language maintenance and revival, and sustainable bilingualism, as long as the philosophy of intercultural education suffuses the communities in which 14 14 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics bilingual education would apply. despite the discouraging state of language transmission in the yungas, the students at the uac are equipped to create the change at the grassroots level that the reform assures from the macro level. 4.2.2 aymara in the public sphere when the students at the uac discussed the underlying shame and embarrassment about using aymara in public, they each asserted that they, individually, had no fear or embarrassment when using the language. they were proud of their culture and their people, and when a situation would arise where they had the opportunity to speak aymara, they would not hesitate. shame and embarrassment of aymara use in public was explained as a subliminal attribute; a historical attitude indexed by the language; a collective identity assumed by all indigenous people as part of a cultural and historical legacy. but i did not witness such a suppressed use of aymara in the coroico municipality. as a whole, people may acknowledge that their identity is pinned down by this debilitating assumption. but reducing it to that singular monologic expression does not convey the shifting, interactional linguistic practices of the aymara people in general, nor the resulting identity work produced by those linguistic practices. identity, as a variegated construction that sustains the ability to shift within context and throughout time, reveals that the “indexical associations imposed from the top down by cultural authorities [may create] ideological expectations among speakers and consequently affect linguistic practice” (bucholtz and hall, 2004:10). however, those same indexical associations do not assume the totality of any collective identity. people told me repeatedly that, on the whole, aymara people were ashamed to use the language. but i witnessed something very different. the local radio station broadcasts aymara programming every morning. the atm machine in town instructs its users in both spanish and aymara. the national newspaper publishes a weekly pull-out section in aymara. political meetings often contain snippits of aymara. and public demonstrations contain aymara songs and stories. popular media in aymara, such as the radio program and the newspaper, would not have been created if ethnic revitalization movements had not opened a space for that type of media to exist. circuitously, the presence of radio programming and print media widen the space for ethnic revitalization, knitting together the aymara community within a nation where they may otherwise be geographically isolated. radio and newspaper offers citizens who normally would not hear from each other the opportunity to communicate. radio is a particularly valuable source of communication, because many older generation aymara citizens cannot read or write, and often a portable radio accompanies farmers out into the fields. i asked the director and commentator of the program, manuel, what type of audience the program is directed toward. the program airs at six in the morning 15 15 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) and targets an older audience, a majority demographic that speaks aymara and would be interested in hearing aymara programming. manuel confessed that aymara is strong here among the older generations, they have no shame using the language. but the younger citizens, teenagers and kids, would have no interest listening to aymara radio. he said that people from the campo, or countryside, often request more aymara programming. the presence of the radio program (including advertisements throughout the day in aymara), and the people’s desire for more programming suggests that, despite teaching their children spanish as the first language and hiding aymara from them, they are not interested in losing their language. it remains an important part of their lives and livelihood, of their culture. in local politics, delegates negotiate the situations when they use spanish and the situations when they use aymara with calculated specificity. i interviewed one young politician, lucio, who had just recently been named the general secretary of the central agraria. a prominent position for a twenty-five year old, lucio used aymara in his public speeches more often than anyone else i met. lucio explained that he uses aymara so that he can communicate with everyone; with people who did not speak spanish. but his explanation did not capture the range of his aymara use. in public meetings, when the entire meeting would be conducted in spanish, lucio would give a speech and then end with an aymara expression, usually raising a supportive, rallying cry from the crowd. lucio uses aymara as a political tool to position himself as a leader; to identify with his community through language; to index a social identity. his choice to speak aymara in an environment where everyone can speak or understand spanish signifies that use of the language functions for some social end other than mere communication. in this regard, he linguistically shapes an identity through interaction with his audience, and that identity has nothing to do with embarrassment or shame. in the context of political meetings, use of the language does not mark lucio as backward, ignorant, uneducated, or a second-class citizen. instead, lucio’s use of the language formulates his subjective position as an aymara leader. likewise, when students volunteer to perform songs or speeches in aymara for national holiday celebrations, like at the day of the sea celebration, their public use of aymara does not mark them as lower class or less valuable members of society. instead, the incorporation of the language functions to symbolically incorporate the diversity of bolivians in the singular bolivian goal of regaining sea coast. it suggests the unification of indigenous bolivians and spanish-descendant bolivians toward a common national endeavor. nationalism thus expresses itself through linguistic difference. each of these examples underscores the use of aymara as a resource by its speakers. they capitalize on the available identities that use of the language might suggest, dialogically creating subject positions within the broader social discourse. not limited to a negative collective identity, use of the language 16 16 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics functions in different domains and contexts to achieve a social goal for the speaker. 5. conclusion the broadening domains of aymara public use serves to degrade the negative social stigmas attached to the language, especially when used against the dominant language, spanish. the fact that the two languages appear side by side in many public environments and popular media demonstrates their ability to complement each other, rather than compete. choosing between one or the other to negotiate an identity in the public sphere no longer carries a singular, negative connotation. citizens make language choices in a variety of different public contexts, using the language to achieve a particular social end—and not one that subjugates them further. the shifting social acceptability of aymara in the public sphere denaturalizes the bourgeois norms of political subordination, transforming relations of power among the linguistic minorities in bolivia. cultural revindication and ethnic movements construct new forms of social organization and value, harvesting legitimacy for linguistic minorities in the national context. politics of identity founded from pluralism penetrate the broad social categories that defined indigenous identity in the colonial era. the emergence of democracy and the adoption of neoliberal reforms scaffold the transformation of identity politics. the law of popular participation and active decentralization allow for each community to create the terms of governance in their local lives, and a more powerful voice on the national level. popular participation and decentralization also respect indigenous culture and methods of governance, sanctioning the value of their practice. rather than imposing government from the top-down, the democratic changes within bolivia proffer a network of representation from the grassroots level that maintains dialogue and respect among the diversity of peoples in the nation. without the neoliberal reforms, law of popular participation and decentralization, and the presence of international lending firms that offer indigenous bolivians voice outside the bounds of the nation-state, the education reform and its philosophy of bilingual and intercultural education could not achieve success. the education reform promises to increase the overall quality of education for bolivians, accomplishing that promise through intercultural and bilingual education for the diversity of the bolivian populace. the new pedagogy abandons the philosophy of ‘civilizing’ its indigenous citizens and instead promotes the value of the multivalent cultures and lifeways. by implementing a more constructive methodology in the classroom, teachers can help students actualize their potential to receive the best possible education, which yields in turn a highly educated populace, creating more jobs and fortifying the economic base of the nation. 17 17 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) bilingual and intercultural education has yet to arrive in the coroico municipality. although other aspects of the reform are well underway, bilingual and intercultural education require the most material production and training programs. teacher training at the uac currently undertakes the feats of producing bilingual materials and bilingual instruction for teachers. as years progress, more and more bilingually trained teachers will be available for local schools, so the bilingual and intercultural modalities can create lasting changes in rural schools. despite first-generation language shift, people in coroico say that aymara should not die out, that it should be preserved. this contradiction appears to be motivated by the fact that parents feel aymara will not serve their children in any domain of life other than the home and for speaking to grandparents. as aymara becomes used and valued more and more in daily life—as radio programming, public demonstrations, and print media have exemplified—then perhaps the next generation will teach aymara to their children, and bilingualism in the coroico municipality will become a sustainable reality. speculations aside, the current gap between policy from the macro level and practice at the grassroots level is wide. macro level policy changes mean nothing unless they are successfully carried out at the grassroots level, which requires incentive and involvement from the people in the communities. if the communities truly do not wish to lose their native language, they command the power to change the trajectory of language loss. language use in the public sphere has been used as a resource; bilingualism an effective tool for creating and negotiating identities within the wider social milieu. identity politics in coroico and bolivia are changing, redefining historical social strata and linguistic diglossia. power wielded from below (not granted from above) illustrates the means by which indigenous communities can alter subjective positions within the nation. a more in-depth study on the efficacy of the education reform in coroico would be a valuable addition to the topics explored in this article. the region is in a dynamic state of transition, and future studies of whether (and how) bilingual and intercultural education improves the quality of life for its speakers in the coroico municipality would further illuminate the advantages of a language-asresource orientation, the viability of linguistic diversity in the globalized world, and the social value of cultural pluralism. the language politics and linguistic situation in coroico sustains as many complexities and contradictions as there are people who live there. the purpose of this article was not to wrap each one up in a neat little package for simplicity of understanding. rather, i hope to have highlighted the realm of potential and possibility that lies below the contradictions and complexities—a realm that illustrates the litigious and dynamic process of democracy and social change within the context of globalization. 18 18 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 education reform and language politics references álbo, xavier. 1999. iguales aunque diferentes. la paz: cipca. -------. 2003. niños alegres, libres y expresivos: la audacia de la educación intercultural bilingue en bolivia. la paz: cipca and unicef. anderson, benedict. 1983. imagined communities. verso: new york. anderson, jeffrey. 1998. “ethnolinguistic dimensions of northern arapahoe language shift.” in anthropological linguistics, vol. 40: 43-108 appadurai, arjun. 1996. “disjuncture and difference in the global cultural economy.” in modernity at large: cultural dimensions of globalization, pp. 27-47. minneapolis: university of minnesota press. bourdieu, pierre. 1991. “the production and reproduction of legitimate language.” in language and symbolic power, pp. 43-65. cambridge: harvard. -------. 1977. outline of a theory of practice. cambridge, uk: cambridge university press. brysk, alison. 2000. from tribal village to global village: indian rights and international relations in latin america. stanford: stanford university press. bucholtz, mary and kira hall (forthcoming). “identity and interaction: a sociocultural linguistic approach.” in alessandro duranti, ed., discourse studies. contreras, manuel e. and maria luisa talavera simoni. 2003. “the bolivian education reform 1992-2002: case studies in large-scale education reform”. washington d.c.: the world bank. freeman, rebecca d. 1996. “dual-language planning at oyster bilingual school: ‘it’s much more than language.’” in tesol quarterly, vol. 30:557-82. freire, paolo. 1993. pedagogy of the oppressed, pp. 43-69. new york: continuum. giddens, anthony. 1990. “the globalising of modernity.” in the consequences of modernity, pp. 63-78. stanford: stanford university press. hannerz, ulf. 1992. “the global ecumene.” in cultural complexity: studies in the social organization of meaning, pp. 217-67. new york: columbia university press. healy, kevin. 2001. llamas, weavings, and organic chocolate: multicultural grassroots development in the andes and amazon of bolivia. notre dame: university of notre dame press. heller, monica. 1999. linguistic minorities and modernity. new york: addison wesley longman inc. 19 19 stockton: education reform and language politics in the coroico municipality of the nor yungas of bolivia published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 20 hornberger, nancy. 1998. “language policy, language education, language rights: indigenous, immigrants and international perspectives.” in language and society, vol. 27: 439-58. irvine, judith t. and susan gal. 2000. “language ideology and linguistic differentiation.” in kroskrity, ed., regimes of language: ideologies, polities, and identities, pp. 35-83. lederbur, kathryn. “popular protest brings down the government.” in special update: bolivia. washington office on latin america. november, 2003. lewellen, ted c. 2002. the anthropology of globalization. westport: bergin and garvey. luykx, aurolyn. 1999. the citizen factory: schooling and cultural production in bolivia. albany: state university of new york press. mignolo, walter d. 2002. “globalization, civilization processes, and the relocation of languages and cultures.” in the anthropology of globalization, pp. 32-51. phillipson, robert. 1992. linguistic imperialism. oxford: oxford university press. quispe, esteban, “los padres de familia quieren logros y no palabras,” la razón: la eib en bolivia. february 2004: suplemento bimensual: 2(3). schiffman, harold. 1996. linguistic culture and language policy. london: routledge. skutnabb-kangas, tove. 2003. “linguistic diversity and biodiversity: the threat from killer languages.” in christian mair, ed. the politics of english as a world language: new horizons in postcolonial cultural studies, pp. 31-52. 20 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/2 doi: https://doi.org/10.25810/v7tn-s780 colorado research in linguistics 6-2005 education reform and language politics in the coroico municipality of the nor yungas of bolivia victoria stockton recommended citation introduction language shift in coroico globalization and bolivia’s democratic reforms changes in bolivian language ideology conclusion teachers╞ differential treatment of culturally and linguistically diverse students during sharing time colorado research in linguistics. june 2008. vol. 21. boulder: university of colorado. © 2008 by laura méndez barletta. teachers’ differential treatment of culturally and linguistically diverse students during sharing time laura méndez barletta university of colorado at boulder this synthesis includes 19 studies that investigate children’s narrative styles during “sharing time.” it looks at teachers’ responses to children’s talk and how teachers’ responses affect children’s school performance and evaluation. findings reveal that when there is a match between the language of the teacher and the student during sharing time, the student receives positive feedback and is allowed to practice her or his oral preparation for literacy. on the other hand, when there is a mismatch between the language of the teacher and that of the student during sharing time, teachers often fail to see the point of what the student is saying. in many cases, the teacher cuts off or interrupts the student, inhibiting the student’s acquisition of literacy skills. this article discusses the differential treatment students receive during sharing time depending upon whether a match or mismatch of teacher/student discourse is present. 1. introduction in many preschool and elementary classrooms across the united states, there is a time of day during which children have the opportunity to share with the rest of the class a narrative about an object brought from home or to give a narrative account about some recent personal experience (michaels 1990). these classroom narrative events are referred to as “sharing time” (also “show and tell,” “rug time,” “news time,” and “circle time” in some classrooms) and are usually centered around the acquisition of literacy (cazden 1985). literacy acquisition is generally focused upon during sharing time through teachers’ questions and comments, especially those aimed at helping students to structure their own discourse (michaels 1981; poveda 2001). sharing time is characterized by faceto-face exchanges between the teacher and students in which children are provided an opportunity to create their own oral texts (cazden 1985), usually by the teacher inviting them to share a narrative of personal experience about their out-of-school lives (cazden 1988). sharing time in preschool and elementary classrooms is of interest because it is typically the only time during the day in which children have the opportunity (during classroom time) to create their own oral texts (cazden 1985). in other words, this is the only time that students are allowed and encouraged to talk freely about a personal experience. sharing time also allows students to talk about their experiences outside of school. during sharing time, students are called on one at a time by the teacher to go to the front of the class (in most cases, they stand next to the teacher, who is seated on a chair) and asked to create a monologue. this is usually followed by a 1 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 2 dialogic exchange between the teacher and the student. through questions and directives, the teacher determines who talks, how long the student talks, and what general or specific topic is addressed. also through questions, comments, and suggestions, the teacher seeks to expand, clarify, or alter the text – in accordance with the teacher’s own, often implicit, expectations about what counts as an appropriate or successful text (michaels 1990). at the same time, teachers can provide support and assistance to the child by expanding on a topic (michaels 1984). in u.s. schools, there are several restrictions or rules (varying from classroom to classroom) that are prevalent during sharing time. for example, teachers may ask students to: 1) talk about one thing; 2) talk about “important” things; 3) not share private family matters; and 4) not talk about television or movies (michaels 1981, 1986). interestingly, poveda (2001) did not find these (or similar) rules in the public school kindergarten classroom that he observed in madrid, spain. on the contrary, poveda found that children’s presentations of oral narratives in spanish classrooms focused on family problems, movies, television shows, and video games (something that is typically not allowed in u.s. schools). research indicates that children from different racial, ethnic, class, and cultural backgrounds bring to the classroom different styles for organizing narratives (labov 1972). some children employ a narrative style that is closer to their home environment, where verbal exchanges take place with familiar people and on a regular basis (hicks 1990). at home, a discourse style that relies on shared background knowledge and assumptions, contextual information, nonverbal cues, and prosody for supplying parts of the intended message is often present (michaels 1983). research has shown that middle-class households with highly literate parents tend to socialize their children into structured routines and patterns of interaction (heath 1982; scollon & scollon 1982; ninio & bruner 1978). through the topic statement/question/answer exchange, the child learns to produce a single, expanded message. linguistic minority children from low socioeconomic backgrounds often are at a disadvantage in american classrooms due to having very little or no access to similar early learning opportunities at home. as a result, children who fail to acquire literacy skills are quite often working-class, minority children from backgrounds that are ethnically and linguistically different from the dominant culture of the school (michaels 1983). in learning to become literate, children have to learn to shift from their homebased conversational discourse strategies to the written language strategies needed to communicate to an unknown audience (michaels 1983; collins & michaels 1986). in other words, they must acquire a new discourse strategy. for example, cazden and john (1968) argue that the “styles of learning” into which native american children are socialized at home greatly differ from those to which they are introduced in the classroom. hymes (1967) indicates that this may lead to sociolinguistic interference when teacher and student do not recognize these differences in their efforts to communicate with one another. 2 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 3 in order to be considered competent, children must conform to the teacher’s implicit expectations as to how information should be organized and presented (michaels 1984). if teachers cannot hear the structure or logic in a student’s story, teachers are generally inclined to assume that no structure exists, that the talk is rambling, unplanned, or incoherent (michaels 1984). this often leads to differential treatment and misevaluation of children. it is important to note that ethnic differences in discourse style have a significant influence in classroom interaction and learning. according to michaels (1981), this problem is not due to racism but rather to differences in ethnic and communicative background, often leading to unintentional mismatches in conversational style. shuy (1981) argues that the language of the classroom is one out of many possible daily language styles. for example, classroom language tends to be different from the language of the home, the playground, and the street. at sharing time, some children’s ways of sharing stories match the expectations of teachers better than others, thereby making it easier for teachers to understand their narrative accounts. when a student’s narrative style matches the teacher’s own style and expectations, collaboration is considered successful, and allows for the improvement of the student’s literacy skills. for such students, this speech event can be considered a preparation for oral literacy. on the other hand, when there is a mismatch between a student’s narrative style and the teacher’s own style and expectations, collaboration is often unsuccessful. here, the student is interrupted or simply misinterpreted. in the long run, such interactions may negatively affect a child’s school performance and evaluation. further, it may exclude the student from the instruction and practice needed to acquire literate discourse strategies. mismatches in student/teacher discourse may result in differential amounts of practice doing literate-style accounting for african american and white children (heath 1982). therefore, it is possible to argue that a student’s narrative style during sharing time can have far-reaching consequences. that is, a child’s narrative style (which always exists in relation to the “preferred style” sanctioned by the teacher and the school system itself) during sharing time can either provide or deny access to key literacy-related experiences depending on the way in which a teacher and child start “sharing” a set of discourse conventions and interpretive strategies (michaels 1981). past research on sharing time reveals that white and african american children’s sharing styles vary considerably. for example, the discourse of white children tends to be tightly organized, centering on a single, identifiable topic. michaels (1981, 1984) calls this discourse style topic-centered. a discourse style that is topic-centered is one that closely matches the teacher’s own discourse style as well as notions of what is considered good sharing. here the teacher and student have a shared sense of what the topic is and are able to collaborate. further, this style gives the teacher an opportunity to build on the student’s contributions and help her or him produce a more focused and lexically explicit discourse (michaels 1981). 3 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 4 in contrast to a topic-centered approach, african american children are more likely to use a topic-associating style. according to michaels (1981), this discourse style consists of a series of implicitly associated personal anecdotes and is generally characterized by an absence of lexicalized connectives other than “and” relating the anecdotes, and no explicit statement of an overall theme or point. this style often gives the impression of the narrative having no beginning, middle, or end, and ultimately, no point at all. the result is often that children seem to ramble on. here, the teacher might have difficulty discerning the topic of discourse and predicting where the talk is going (michaels 1981, 1983). the teacher’s questions may also be mistimed, stopping the student at mid-clause as well as interrupting the student’s train of thought. in addition, the teacher may not build fully on the child’s own narrative intentions. this article summarizes and synthesizes the studies that have been done on children’s narrative styles during sharing time, teachers’ responses to children’s talk, and how teachers’ responses affect the talk as well as the children’s school performance. implications of these studies for practice are also discussed. to date, no other syntheses have been published summarizing this body of research. 2. method 2.1 selection of studies taking the approach used by klingner and vaughn (1999) in their synthesis of student perceptions of instructional procedures, the studies presented in this synthesis were selected based on a two-step procedure. a thorough search on “sharing time” was conducted in order to ensure that all of the existing publications in this area were located. in order to gather as much information as possible on sharing time, four modes of searching were used in this synthesis: (a) searches in subject indexes, (b) citation searches, (c) consultation, and (d) browsing. step 1: initial selection of studies searches in subject indexes. similar to klingner and vaughn’s (1999) synthesis, i conducted computer searches through two databases to identify relevant published articles, papers presented at major national educational conferences, final reports, and dissertations. these searches consisted of utilizing the educational resources information center (eric) and proquest digital dissertations. a number of computer searches through eric database advanced search were conducted (not limited by date) and included sets of descriptors such as: “sharing time and teacher role and racial differences” and “sharing time and child language and teacher response.” when the first set of studies was identified using the above descriptors, the major and minor descriptors found in these 4 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 5 articles were examined in order to find other articles. a second stage of searches was then conducted with the following sets of descriptors: “sharing time and oral language,” “story telling and cultural differences,” “sharing time and teacher student relationship,” and “show and tell and classroom environment.” once additional studies had been identified using these descriptors, more stages of searches were initiated. additional searches were conducted of proquest digital dissertations to gather information not available through other sources. i used various types of descriptors in this database such as “sharing time and classroom discourse,” “circle time and classroom discourse,” and “classroom discourse and literacy.” abstracts that matched these descriptors were reviewed to determine if they included a focus on children’s narrative style during sharing time, teachers’ response and interaction with the children during sharing time, and how the teachers’ response and interaction affects children’s learning. citation searches. lists of citations were checked from relevant studies to assure that every article cited was looked at for possible inclusion in the sample. this approach was helpful due to the identification of articles that might not have appeared in eric or in the dissertation abstract database. consultation. i attempted to locate articles that might be “in press” or “in progress” by contacting several researchers who have published articles on sharing time in the past. i sent letters to several researchers asking if they had any articles on sharing time that were “in press,” “in progress,” and/or if they were aware of any other researchers who had written articles focused on teachers’ response/interaction with students during sharing time. browsing. hand-searches of the following journals were conducted: linguistics and education, journal of education, anthropology and education quarterly, language arts, bilingual research journal, and theory into practice. i browsed through the journals’ table of contents, going back twenty-five years. these articles were chosen because researchers writing on sharing time published their articles in these journals. browsing through these articles allowed me to search for articles that were not located in the eric database. step 2: final selection of studies in order for a study to be included in this synthesis, it must contain data on students’ narrative style during sharing time, the teachers’ response/interaction with the students during sharing time, and teacher/student collaboration during sharing time. i did not include articles that focused on teachers’ attitudes about students’ narrative during classroom speech events. further, i did not include studies about why some students are less talkative than others during classroom speech events nor studies on how students learn to participate in sharing time. when a study included multiple components, i included only relevant components that fit my criteria. 5 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 6 2.2 analysis procedures after i assembled the set of articles, my next step was to read and code them using the following categories: purpose of the study, participants’ narrative during sharing time, and applicable findings. when studies had multiple purposes, including measures and results, only those that pertained to this synthesis were included in my analysis. for example, with the davis and golden (1994) study that examined teachers’ perceptions about children not attending school with an increased english verbal communication ability as well as how misunderstandings become interpreted by teachers, i only included information about the misunderstandings; i omitted the teacher’s perceptions because it did not address the purpose and criteria for this synthesis. in order to summarize the findings of the articles i collected, i read the articles and recorded their descriptions and key findings in a database. i then pulled out common themes from the articles and transferred them into another database. finally, i re-read the articles to determine whether the findings should be included or whether the findings were unrelated to the purpose of this synthesis. 3. results 3.1 participants the studies in this synthesis included participants in kindergarten through seventh grade (one study included teachers as participants analyzing students’ narratives). it was challenging to determine the exact number of participants in each study due to the fact that not all of the studies under consideration provided the total number of participants. for example, some studies included only the number of classrooms observed (for example, danielewicz et al. (1996) studied one first-grade classroom). of the 19 studies in this synthesis, eight studies included participants of different ethnic backgrounds (other than african american and white), two studies included an equal number of african american and white participants, four studies included only african american participants, in one study participants were predominately african american, and in four studies the ethnic background of participants was not reported. table 1 provides a summary of the studies included in this synthesis. the numbers assigned to the studies will be used to refer to them throughout the article. 6 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 7 table 1 summary of the 19 studies study purpose participants 1. cunningham, 1976-77 investigate teachers’ tendency to correct blackdialect-specific miscues and their ability to recognize black dialect 214 teachers analyzing work done by students in grade 4 2. danielewicz, rogers, & noblit, 1996 investigate students’ narrative style and interaction patterns during sharing time (teacher-led format and child-led format) one first-grade classroom 3. daniell, 1996 expand on michaels’ (1986) findings of deena’s story during sharing time (and offers a critique to a student’s narrative style) one african american female student in the first grade 4. davis & golden, 1994 discuss the differences between students’ and teachers’ communication style (and interpretation of utterances) 300 kindergartners (98% african american) 5. gallas, 1992 present one student’s narrative style and looks at the social nature of the classroom community. provides information on how children and teachers can work together to understand each other’s stories first-grade classroom of 22 students (3 african american, 11 white, 6 japanese, 1 south african, and 1 ethiopian) 6. gee, 1985 give an analysis of an african american student’s narrative style and discusses the teacher’s response to the narrative. also seeks to explain how the child makes sense of her experiences through narrative one 7-year-old african american female student 7 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 8 table 1 (continued) summary of the 19 studies study purpose participants 7. gee, 1989 offer an analysis of an african american and white student’s narrative style during sharing time one 11-year-old african american female student and one 11-year-old white female student 8. hyon & sulzby, 1994 discuss students’ narrative style during sharing time forty-eight african american low-income kindergartners 9. mccabe, 1997 synthesizes research on the importance of stories in classrooms and how students’ narrative style differs from culture to culture synthesis 10. michaels, 1981 african american and white students’ narrative style is analyzed (and the teacher’s response to their discourse) one first-grade classroom of diverse students 11. michaels, 1983 discuss the significance of ethnic differences in students’ narrative style and its influence on classroom interaction and learning; looks at teacher/child collaboration at sharing time four boston integrated classrooms (1st, 2nd, and two combined 1st-2nd grades) and one berkeley classroom (half middleclass white and half working-class african american) 12. michaels, 1984 look at students’ narrative style during sharing time (also teacher/child collaboration) one 2nd-grade classroom of ethnically diverse students 8 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 9 table 1 (continued) summary of the 19 studies study purpose participants 13. michaels & collins, 1984 discuss situations in which sharing turns result in more successful teacher/child collaboration and extended discourse than others (including students’ narrative style) one 1st-grade integrated classroom 14. michaels & foster, 1985 look at students’ narrative style (and teacher’s response) to or evaluation of children’s discourse one combined 1st-2nd grade classroom of 20 ethnically diverse students 15. michaels, 1986 discuss teacher’s response to students’ narrative style one 1st-grade classroom (half white and half african american students) 16. michaels & cazden, 1986 discuss how discourse patterns related to ethnic background affect the quality of teacher/child collaboration one 1st grade classroom (30 students) of ethnically diverse students (14 white, 15 african american, 1 asian) 17. michaels, 1990 discuss teacher’s response to students’ narrative style sharing time in one 1st grade classroom and one 2nd-grade classroom and a composition writing activity in one 6th-grade classroom 18. poveda, 2001 look at similarities and differences between sharing time speech events in spain and u.s. schools one kindergarten classroom (18 students) of ethnically diverse students (spanish gypsies, african and latin american immigrants, and non spanish gypsies) 19. puro & bloome, 1987 look at teacher/child interaction patterns one kindergarten class, one 1st-grade reading group, and one 7th grade classroom 9 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 10 3.2 data sources the primary data sources for all of these studies were students’ narrative accounts and teachers’ responses during sharing time. measures included: observations of teacher-led sharing time (2, 4, 5, 10, 11, 12, 13, 15, 16, 17, 18, 19), observations of student-led sharing time (2, 14), students’ language and interaction patterns (1, 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19), teachers’ responses to students’ narratives (2, 4, 5, 10, 11, 13, 15, 16, 17, 18), written narratives (13, 17), oral interviews (4, 8, 13, 15, 17), and written questionnaires (1). 3.3 description of studies of the 19 studies that met the criteria for inclusion in this synthesis, 14 were published in refereed journals and five were published as book chapters. all studies reported that one of their purposes was to investigate students’ language style during sharing time. in addition, all studies sought to look at teachers’ responses to students’ language style as well as collaborative exchanges between students and their teachers. 3.4 summary of findings an analysis of the articles generated eight categories of findings: students’ narrative style, teacher response, teacher/student collaboration, interpretation of utterances, interaction patterns, students’ and teachers’ communication style, teachers’ tendency to correct black dialect-specific miscues, and similarities and differences between sharing time speech events in spain and u.s. schools. some of the studies addressed only one of these categories, while others overlapped and covered multiple categories. the applicable categories addressed by each article are highlighted in bold text in the purpose statements listed in table 1. 3.4.1 student’s narrative style fourteen studies address students’ narrative style during sharing time in some way. students’ narrative style during sharing time was the primary focus of 12 studies (2, 5, 6, 7, 8, 9, 10, 11, 12, 14, 15, 17), and a secondary focus of two (3, 13). one theme related to students’ narrative style was identified: topic-centered style and topic-associating style. nine studies address the two styles of narratives that children use during sharing time: topic-centered and topic-associating (3, 6, 8, 10, 11, 12, 13, 15, 17). topic-centered narratives are characterized as discourses that are tightly organized around a single topic. in contrast, topic-associating narratives are characterized by frequent shifts in time, location, and key characters. research on 10 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 11 sharing-time stories indicate that white students tell topic-centered stories, while a significant number of stories told by african american students are topicassociating stories (michaels 1981; michaels & cazden 1986). in one study, an african american female student’s (deena) narrative style is analyzed and identified as topic-associating (3). according to deena’s teacher, her story appears to be a stitching-together of unrelated pieces of information (daniell 1996). as a result, deena’s teacher was not successful in helping her to structure and clarify her narrative. deena’s topic-associating story caused her teacher to ask questions at inappropriate times, causing deena to lose her train of thought. in another study, the narrative of a seven-year-old african american girl (“l”) is examined (6). “l’s” topic-associating narrative style during sharing time is not immediately recognizable by her teacher and as a result appears as incoherent. in the end, “l” is given less instructional time and attention than those children who use a topic-centered style. gee (1985) argues that “l” is considered to be a master of making sense of her experience, and she carries her story with full utilization of prosody, time and sequence markers, parallelism, and repetition. the purpose of another study (8) was to assess the frequency of topicassociating narratives among african american kindergartners. the study found that the stories told by the participants included both topic-centered and topicassociating narratives. results revealed that out of the 48 narratives, there were 16 topic-associating stories, 28 topic-centered stories, and 4 stories whose category was not clear. further, results indicate that the topic-associating style was not predominant among african americans, as michaels’ observations have shown. another study (11) presents a pair of excerpts (including both topicassociating and topic-centered styles) with the teacher’s response to the students’ discourse style. the study examines the discursive skills that are required in literate-style communication and focuses specifically on the effect that differences in discourse style may have on teacher/student collaboration. the author presents findings that indicate that middle-class, highly literate parents engage their children in structured routines and patterns of interaction (contrary to workingclass, less literate parents). this practice allows children to develop their communicative abilities, thereby preparing children for the demands of literate discourse. further, the author argues that a child’s use of a discourse style that is at variance with the teacher’s expectations decreases the quality of instruction in key classroom activities, which then interferes with the child’s development of a prose-like discourse style. another study (12) examines children’s preferred strategies for structuring a narrative account. findings indicate that 96 percent of white children, 34 percent of african american males, and 27 percent of african american females used a topic-centered narrative style during sharing time. michaels suggests that african american children are more likely to tell narratives using a topicassociating style while white children use a topic-centered style. further, she argues that teachers are better able to follow cues in topic-centered discourse due 11 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 12 to turns meeting the teachers’ expectations about where certain information should be located and how a topic should be developed. in a number of studies (10, 13, 15, 17), topic-centered and topic-associating narratives are illustrated and analyzed. the topic-associating narrative in all these studies is that of deena, an african american female. during deena’s sharing turn, her teacher repeatedly tells her to talk about “one thing,” interrupting and questioning her several times. these studies indicate that such types of responses/actions by teachers (with topic-associating children) interfere with students’ train of thought, causing them to stop talking or revert to one or twoword responses. when deena’s teacher was asked what she thought about topicassociating turns, she explained that students really don’t think about what they want to say in advance and simply talk off the top of their heads. these studies indicate that deena and her teacher were working within their own sharing time schema; that is, without a shared sense of topic as well as a shared set of discourse conventions. on the other hand, these studies indicate that students who used a topic-centered narrative were understood by the teacher; the teacher was successful at picking up on the students’ topic. the teacher’s questions occurred after pauses, descended from general to specific, and the teacher’s responses and clarifications built on students’ own contributions. 3.4.2 teacher response five studies focused on teacher response to students’ narrative style during sharing time (6, 10, 14, 15, 17). all five studies emphasized teachers’ responses when students used a topic-centered narrative style versus those that used a topicassociating narrative style. teachers tended to offer students who used a topiccentered style a scaffold on which to build their narratives. for example, through statements, questions, and responses, teachers were able to elicit more explicit information on the students’ topic. students in these studies that used a topiccentered narrative style received interactive support from teachers as well as extended practice for learning the narrative demands of the classroom. on the other hand, teachers were less successful at providing a scaffold for students who used a topic-associating narrative style. for example, during their narrative accounts, teachers’ questions were often mistimed, teachers interrupted students at mid-clause, and students’ turns were often cut short by the teacher. this tended to throw the children off balance and ultimately interrupt their train of thought. 3.4.3 teacher/student collaboration four studies focused on teacher/student collaboration during sharing time (11, 12, 13, 16). in one study (11), the teacher actively participated, asking questions and making comments to help students clarify, structure, and expand their discourse. in this study, students were encouraged to be clear and precise, and to put all the information their audience needed into words rather than relying on 12 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 13 shared background knowledge or contextual cues to communicate part of the intended message (michaels 1983). michaels’ study found that the teacher collaborated more successfully with some students than with others; according to michaels, this collaboration depended on the degree to which the teacher and student shared a set of discourse conventions. it was revealed that the berkeley teacher frequently used a confrontational strategy with topic-associating students telling them to talk about “important things” or “one thing only.” the boston teachers, in contrast, rarely used overtly confrontational strategies (michaels 1983). finally, michaels found that problems in teacher/student collaboration stem from a mismatch between a teacher’s and student’s narrative strategy and use of prosody. further, michaels believes that these mismatches, over time, result in differential amounts of practice and instruction for children in organizing information according to a literate model. in a second study (12), an interactional pattern (“vertical construction”) that can result in collaborative development of a topic is described. through this statement/question/answer exchange, the teacher and student collaborate to produce a single, expanded message. this study focused on the role that a secondgrade teacher played during sharing time. it was found that the teacher played a pivotal role as listener and responder, addressing questions and comments to the child sharing or the audience at large, trying to help the child clarify and expand his or her discourse (michaels 1984). while african american and white children in this study used a sharing intonation (i.e., up-talking) strategically, the teacher was better able to follow these cues in topic-centered discourse because these turns met her expectations about where certain information should be located and how a topic should be developed (michaels 1984). in the third study (13), collaborative exchanges between a teacher and her students at sharing time were analyzed. it was found that some sharing turns resulted in more successful teacher/child collaboration and extended discourse than others. as a result, some children seemed to get more practice using literate discourse strategies than did others (michaels & collins 1984). it was concluded that the teacher/child interaction was asynchronously paced when students used an “oral discourse style” during sharing time (as opposed to a “literate discourse style”). when students used an oral discourse style, the teacher made frequent interruptions, thematically inappropriate comments, and, as a result, there was minimal collaboration between the teacher and her students. michaels and collins argue that lack of teacher/student collaboration results in a pattern of differential treatment and negative evaluations that diminish students’ access to the kind of instruction and practice necessary for the acquisition of literacy. the final study in this group (16) focuses on teacher/child collaborative exchanges during sharing time. this study pays attention to the following pattern: the student says something (often in response to a teacher’s question), is again queried by the teacher, and then provides more information as elaboration. it is argued that through the above sequence of questions and answers, the teacher and student construct (together) a single, expanded message. furthermore, this kind of 13 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 14 exchange gives the student practice at being lexically explicit. the teacher in this study participated actively at sharing time, and sharing time in this classroom was considered a kind of “oral preparation for literacy.” however, it was found that not all of the students gained equal access to help. the teacher collaborated more successfully with some children than with others at sharing time, depending on the degree to which the teacher and student started out sharing a set of discourse conventions. michaels and cazden found that collaboration stemmed from a match between teacher and student’s narrative strategies and use of prosody. 3.4.4 interpretation of utterances one study (4) focused on teachers’ interpretation of students’ utterances in a kindergarten center. in this study, two teachers explain that the lack of teacherstudent communication is due to some children not having much language experience and therefore, not having sufficient vocabulary. they explain that many students lack the kinds of experiences at home that would help them prepare for school. while these teachers interpret some students’ verbal communication in terms of a deficiency, they are unable to see other possible explanations for lack of student participation, such as miscommunication due to differences in interactional styles. both teachers believe that if a student’s behavior does not match the school language and expectations (that is, mainstream language), then that student comes from a home lacking in language and “proper” ways of behaving (davis & golden 1994). davis and golden argue that the ways in which these teachers evaluate students and engage in classroom interaction can be harmful to the students with whom they work. 3.4.5 interaction patterns two studies (2, 19) focused on student interaction patterns during sharing time. one of these studies (2) investigated students’ language and interaction patterns during teacher-led and a student-led sharing time events. during the teacher-led speech event, students spoke the language of the school modeled by the teacher, responded to the teacher’s script, and spoke the words insisted upon by the teacher. in the teacher-led event, the teacher controlled the conversation and steered the conversation toward categories of acceptable talk. during the student-led speech event, students in the role of the sharer established the topic and controlled the interaction to extend discussion. danielewicz et al. (1996) found that this student-led dialogue fostered peer culture and a sense of individual and group identity. further, they found that student-led sharing speech events allow students to gain power and control while simultaneously building community through shared discussions and common rituals. the purpose of the second study (19) was to help students acquire reading vocabulary as well as develop their reading comprehension skills. when a student answered a question in a way not acceptable to the teacher, the latter modeled 14 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 15 how the former was to structure the response (that is, in a book-like sentence). here, students learned how to formulate an appropriate answer to a teacher’s question as well as how to construct a book-like sentence. puro and bloome (1987) argue that teachers and students interpret each other’s messages in terms of the interactional context. this can be observed, for example, when a student has to reformulate his/her answer in terms of a book-like sentence. from the interactional context in this event, the student and the others in the group learn how to structure their relationship to printed text and what constitutes comprehension (puro & bloome 1987). 3.4.6 students’ and teachers’ communication styles one study (4) focused on student and teacher communication styles during storybook reading time. the study suggests that the ways in which teachers engage in classroom interaction can be harmful to the children they work with. during storybook reading time, when students answered teachers' questions in unison, two teachers allowed it in some instances but not in others. these two teachers reinforced “appropriate” behavior during storybook reading by either ignoring the students or asking a specific student if s/he wanted a talking turn. a third teacher utilized several different strategies with the expectation that children would respond in unison. for example, one strategy was to pause and have children complete the teacher’s text. if the students responded correctly, the teacher affirmed this by restating their response. 3.4.7 teachers’ tendency to correct black dialect-specific miscues one study (1) focused on teachers’ tendency to correct black dialect. this study investigated teachers’ attitudes toward non-meaning-changing miscues and to see if these attitudes were different for black-dialect-specific miscues. it also aimed to discover if a relationship existed between the number of black-dialect miscues teachers indicated they would correct and the number of speech samples they recognized as being spoken mostly by african american students. teachers’ responses indicated that they would correct significantly more black-dialectspecific miscues (78 percent “would correct” responses) than non-dialect-specific miscues (27 percent “would correct” responses). cunningham (1976-77) argues that a major obstacle to reading success for african american children may be found not in their language but in their teachers’ attitude toward and reaction to that language. 3.4.8 sharing time speech events in spain and u.s. schools one study (18) compared differences between sharing time in spain and the u.s. poveda (2001) states that “la ronda” (the round) and sharing time are 15 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 16 distinguishable in participation structures, conversational topics, children’s initiations, and teachers’ feedback. in spain, la ronda is an event used to socialize students into a classroom community that shares a number of behavioral, affective, and cognitive standpoints. sharing time in the u.s., on the contrary, is an instructional event in which certain linguistic-discursive forms, often explicitly related to later literacy development, are practiced (poveda, 2001). 4. discussion sharing time is an activity that gives students the opportunity to share with the rest of the class a narrative about an object brought from home or to give a narrative account of some recent personal experience. in u.s. schools, these narrative events are usually seen as a type of oral preparation for literacy, focusing on academic skills and content (harris & fuqua 2000). a number of studies have suggested that children from different racial, ethnic, and cultural backgrounds attend school with different skills for giving narrative accounts. studies in this synthesis indicate that when a student’s discourse style matches the teacher’s own style and expectations, collaboration is synchronized and allows for informal practice and instruction in the development of a literate discourse style. in contrast, when the student’s narrative style is inconsistent with the teacher’s expectations, collaboration is often unsuccessful and, over time, may adversely affect school performance and evaluation. the present study summarizes 19 studies conducted within the last 28 years; its goal is to come to a better understanding of students’ narrative styles during sharing time and teachers’ responses to them. further, it seeks to document the differential treatment students receive depending upon their narrative style. studies addressed the following aspects of students’ narrative style: teachers’ response/feedback, teacher/student collaboration, interpretation of students’ utterances, teacher/student interaction patterns, students’ and teachers’ communication style, and students’ school performance and evaluation. 4.1 implications for practice this synthesis provides direct implications for school teachers and administrators. one of the most significant findings in this synthesis is the suggestion that teachers view the majority of african american students’ language (during sharing time) as uncommunicative and unacceptable. furthermore, teachers interpret many african american students’ communicative style in terms of a cognitive handicap and view them as coming from a home lacking in language. teachers in these studies tend to label the majority of african american children’s narratives as having no beginning, middle, and end, and ultimately, no point at all. in the end, teachers respond differently to african american vernacular english (topic-associating style) and standard american english (topic-centered style). 16 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 17 teachers’ attitudes of student’s language style during sharing time often lead to children having differential access to learning opportunities in the classroom. most studies in this synthesis indicate that after an african american’s narrative account there was a complete absence of teacher/student collaboration, something that occurred very infrequently with white students. thus, the discourse style employed by a student influences the kind and amount of teacher/child collaboration that occurs. if a student uses a topic-centered style, the teacher is successful at picking up the child’s topic and offering a scaffold on which to build. in addition, the teacher offers interactive support, asking general to specific questions, thereby building on the child’s own contributions. on the other hand, if a student uses a topic-associating style, the teacher is less successful at providing a scaffold. further, the teacher asks questions that are often inappropriate and thereby mistimed. the result is that the teacher often interrupts the child at midclause and throws her or him off-balance. researchers in this synthesis believe that mismatches in teacher/student discourse frequently result in interruptions, misunderstandings, and misassessment and misevaluation of children’s abilities (e.g., evaluating children as less capable of producing organized, well-planned texts). it is indicated that over time, these negative evaluations may in turn influence the teacher’s expectations and treatment of and attitudes toward these children as learners. michaels’ research (1981, 1983, 1986) argues that mismatches between teachers and students negatively impact the literacy instruction children receive. she goes on to state that these misunderstandings negatively affect the teacher-student relationship, a crucial factor in learning. moreover, collins (1982) suggests that any sort of communicative mismatch between the language of the teacher and student will reinforce decisions about which students will be classified as highability and which will be classified as low-ability learners. he goes on to argue, following anyon (1981), that such decisions can influence allocation of the teacher’s time, compounding the general tendency in public schooling to allocate the smallest percentage of resources to those who need them most (collins 1982). students make meaning of their teachers’ responses toward their own (and others’) way of speaking during sharing time. for example, during an interview conducted by michaels, deena (a six-year-old african american student) expressed a keen sense of frustration about being interrupted during sharing time. she saw being interrupted as an indication that the teacher was simply not interested in what she had to say: sharing time got on my nerves. she was always interruptin’ me saying’, “that’s not important enough,” and i hadn’t hardly started talkin’!. . . i felt like slappin’ her upside the head,. . .sayin’ ‘well it’s important to me, so you just listen when i’m talkin’ to you woman! (michaels 1990) 17 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 18 deena’s older sister recalled similar frustrations from her sharing experiences five years earlier in both kindergarten and first grade. these studies teach us that the problem of mismatched discourse appears to relate more generally to differences in ethnic and communicative backgrounds, leading to unintentional mismatches in conversational style. over time, such mismatches may result in differential amounts of practice doing literate-style narrative accounts for african american and white children in class, which may ultimately affect children’s progress in the acquisition of literacy skills. further, mismatches can greatly influence children’s participation and educational success. in the end, improving teacher/student collaboration can increase students’ opportunities to learn by enhancing students’ access to the kind of quality instruction that they need. these studies also reveal that children’s communicative style is associated with their cultural identity and presentation of self. they suggest that teachers and schools do not understand or value students’ mode of expression, do not see students’ language style connected to a culture and sense of self, and that teachers do not give access to the instruction that would ensure that students could switch narrative style, let alone do so in a way that does not threaten their own sense of self. cazden (1976) argues that in out-of-school conversations, one’s attention (as speakers and listeners) is on the meaning, the intention, of what someone is trying to say. she maintains that teachers have gotten into the habit of hearing with different ears once they enter the classroom; they only hear the errors to be corrected. 4.2 limitations there were three limitations to the way research was conducted in the studies reviewed that deserve mentioning. first, no studies were found that focused on english language learners. there is a need for more research to focus on linguistically (as well as culturally) diverse students during sharing time. second, almost all data in these studies were based on observations of african american and white children. it would be critical to look at different ethnic and cultural backgrounds to learn whether different narrative styles during sharing time are present. third, african american children in the studies tended to be workingclass (or poor) and come from inner cities. it would be important to expand this research to include african american children (as well as children from other ethnic backgrounds) that come from various income brackets as well as geographic regions. there are some questions that remain unanswered after reviewing the studies in this synthesis. first, it was not indicated whether it is possible for similar problems to be present in other contexts where teachers and students attempt to collaborate in the joint development of a coherent message. for example, can mismatches exist between teachers and students during group reading lessons? second, there was no suggestion regarding the frequency of topic-associating 18 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 19 discourse among african american children. for instance, how often and within what contexts does a topic-associating discourse occur? third, while sharing time is seen as an oral preparation for literacy, its influence on children’s reading ability is unclear. it would be valuable to explore how children’s discourse style (topic-centered and topic-associating) affects their reading ability. references anyon, jean. 1981. “social class and school knowledge.” curriculum inquiry 11: 3-42. cazden, courtney b. 1976. “how knowledge about language helps the classroom teacher or does it: a personal account.” the urban review 9(2): 74-90. cazden, courtney b. 1985. “research currents: what is sharing time for?” language arts 62(2): 182-188. cazden, courtney b. 1988. classroom discourse: the language of teaching and learning. portsmouth, n.h.: heinemann educational books. cazden, courtney b. and vera p. john. 1968. “learning in american indian children.” in styles of learning among american indians: an outline for research. washington, d.c.: center for applied linguistics. collins, james. 1982. “discourse style, classroom interaction and differential treatment.” journal of reading behavior 14(4): 429-437. collins, james and sarah michaels. 1986. “speaking and writing: discourse strategies and the acquisition of literacy.” in jenny cook-gumperz (ed.) the social construction of literacy. london: cambridge university press. cunningham, patricia m. 1976-77. “teacher's correction responses to blackdialect miscues which are non-meaning-changing.” reading research quarterly 4: 637-653. danielewicz, jane m., dwight l. rogers, and george w. noblit. 1996. “children's discourse patterns and power relations in teacher-led and child-led sharing time.” qualitative studies in education 9(3): 311-332. daniell, beth. 1996. “deena’s story: the discourse of the other.” jac 16(2): 253-264. davis, kathryn a. and joanne m. golden. 1994. “teacher culture and children’s voices in an urban kindergarten center.” linguistics and education 6: 261287. gallas, karen. 1992. “when the children take the chair: a study of sharing time in a primary classroom.” language arts 69: 172-182. gee, james paul. 1985. “the narrativization of experience in the oral style.” journal of education 167(1): 9-35. gee, james paul. 1989. “two styles of narrative construction and their linguistic and educational implications.” journal of education 171(1): 97115. 19 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 20 harris, teresa t. and j. diane fuqua. 2000. “what goes around comes around: building a community of learners through circle times.” young children 55(1): 44-47. heath, shirley bryce. 1982. “what no bedtime story means: narrative skills at home and at school.” language in society 11(1): 49-77. hicks, deborah. 1990. “kinds of narrative: genre skills among first graders from two communities.” in allyssa mccabe and carole peterson (eds.), developing narrative structure, 55-87. hillsdale, n.j.: lawrence erlbaum. hymes, dell. 1967. “on communicative competence.” in renira huxley and elisabeth ingram (eds.), mechanisms of language development. london: centre for advanced study in the developmental science and ciba foundation. hyon, sunny and elizabeth sulzby. 1994. “african american kindergartners’ spoken narratives: topic associating and topic centered styles.” linguistics and education 6: 121-152. klingner, janette and sharon vaughn. 1999. “students’ perceptions of instruction in inclusion classrooms: implications for students with learning disabilities.” exceptional children 66(1): 23-37. labov, william. 1972. “the transformation of experience in narrative syntax.” in william labov (ed.), language in the inner city: studies in the black english vernacular, 354-396. philadelphia: university of pennsylvania press. mccabe, allyssa. 1997. “cultural background and storytelling: a review and implications for schooling.” the elementary school journal 97(5): 453-473. michaels, sarah. 1981. “‘sharing time’: children’s narrative styles and differential access to literacy.” language society 10: 423-442. michaels, sarah. 1983. “the role of adult assistance in children’s acquisition of literate discourse strategies.” the volta review 85(5): 72-86. michaels, sarah. 1984. “listening and responding: hearing the logic in children’s classroom narratives.” theory into practice 23(3): 218-224. michaels, sarah. 1986. “narrative presentations: an oral preparation for literacy with first graders.” in jenny cook-gumperz (ed.), the social construction of literacy, 94-116. london: cambridge university press. michaels, sarah. 1990. “the dismantling of narrative.” in allyssa mccabe and carole peterson (eds.), developing narrative structure, 303-351. hillsdale, n.j.: lawrence erlbaum. michaels, sarah and courtney b. cazden. 1986. “teacher/child collaboration as oral preparation for literacy.” in bambi b. schieffelin and perry gilmore (eds.), the acquisition of literacy: ethnographic perspectives, 132-154. norwood, n.j.: ablex publishing. michaels, sarah and james collins. 1984. “oral discourse styles: classroom interaction and the acquisition of literacy.” in deborah tannen (ed.), coherence in spoken and written discourse, 219-244. norwood, n.j.: ablex publishing. xii in the series. 20 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/3 doi: https://doi.org/10.25810/tf1r-yj06 teachers' differential treatment of culturally and linguistically diverse students 21 michaels, sarah and michele foster. 1985. “peer-peer learning: evidence from a student-run sharing time.” in angela jaggar and m. trika smith-burke (eds.), observing the language learner, 143-158. newark, del.: international reading association national council of teachers of english. ninio, anat and jerome bruner. 1978. “the achievement and antecedents of labeling.” journal of child language 5: 1-15. poveda, david. 2001. “la ronda in a spanish kindergarten classroom with a cross-cultural comparison to sharing time in the u.s.a.” anthropology & education quarterly 32(3): 301-325. puro, pamela and david bloome. 1987. “understanding classroom communication.” theory into practice 26(1): 26-31. scollon, ron and suzanne b. scollon. 1982. “cooking it up and boiling it down: abstracts in athabaskan children’s story retellings.” in deborah tannen (ed.), spoken and written discourse. norwood, n.j.: ablex publishing. shuy, roger w. 1981. “learning to talk like teachers.” language arts 58(2): 168-174. 21 me?ndez barletta: teachers’ differential treatment of culturally and linguistically diverse students during sharing time published by cu scholar, 2008 colorado research in linguistics 6-2008 teachers’ differential treatment of culturally and linguistically diverse students during sharing time laura méndez barletta recommended citation microsoft word cril_paper_mendez-barletta.doc language acquisition and the 'dative alternation' 1 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 2 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 3 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 4 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 5 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 6 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 7 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 8 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 9 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 10 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 11 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 12 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 13 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 14 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 15 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 16 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 17 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 18 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 19 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 20 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 21 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 22 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 23 shay: language acquisition and the 'dative alternation' published by cu scholar, 1998 24 colorado research in linguistics, vol. 16 [1998] https://scholar.colorado.edu/cril/vol16/iss1/2 doi: https://doi.org/10.25810/mb11-6112 colorado research in linguistics 1998 language acquisition and the 'dative alternation' erin shay recommended citation tmp.1537466529.pdf.8ow9u language ideologies in the arabic diglossia of egypt colorado research in linguistics. june 2010. vol. 22. boulder: university of colorado. © 2010 by susanne stadlbauer. language ideologies in the arabic diglossia of egypt susanne stadlbauer university of colorado at boulder this paper surveys studies on language ideologies in the arabic diglossic environment of present-day egypt. specifically, it discusses linguistic and cultural implications of language ideologies associated with classical arabic (ca), modern standard arabic (msa), egyptian arabic (ea), and english in the cairo area. the language ideologies of these varieties are a product of both the past and the present: they emerged during british colonialism in the late nineteenth century and are maintained in the postcolonial climate through discourses on the purity of classical arabic, on the linguistic corruption of the dialects, and on the increasing use of english as a symbol of western capitalism and modernity. aligning with woolard’s (1998) definition of language ideology as a mediating link between linguistic features and social processes, this study demonstrates how language ideologies are communicated in structural aspects of the language varieties in the arabic diglossia and how egyptians use language varieties strategically to access the symbolic power of these ideologies. it argues that studies of language ideologies, language features, and discursive interaction are inseparable in uncovering how language is used in the arabic diglossia in egypt. 1. introduction studies on language ideologies and language-related historical studies on nationalism in egypt demonstrate that “the equation of language and nation is not a natural fact but rather a historical, ideological construct” (woolard 1998:16). language ideologies are social constructs that are illuminated through a microanalysis of linguistic structures in discourse and a macro-analysis of the factors that lead to asymmetries in how languages are perceived. in line with woolard, i emphasize the analysis of language ideologies within the linguistic practices of specific cultural settings, which means that language ideologies cannot have a single interpretation. they are in dialectical relation with social, discursive, and linguistic practices and often determine “which linguistic features get selected for cultural attention and for social marking, that is, which ones are important and which ones are not” (schieffelin & doucet 1998:285). in short, language ideologies are never about language alone, but rather, envision and enact ties of language to identity, to aesthetics, to morality, and to epistemology. through such linkages, they underpin not only linguistic form and use but also the very notion of the person and the social group, as well as such 1 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 2 fundamental social institutions as religious ritual, child socialization, gender relations, the nation-state, schooling, and law (woolard 1998:3). taking these definitions as a conceptual blueprint, this study shows how these linkages play out in a specific linguistic and cultural setting: arabic diglossia in egypt. present-day language ideologies in egypt find their beginnings in british colonization from 1882 to 1922 (mitchell 1988, suleiman 2003, amongst many). in this period, the colonizers restructured egyptian society according to western ideals of modernity and economic progress. among other projects, they initiated anti-arabic, pro-english language policies that assigned symbolic value to these languages: arabic was depreciated because it was perceived as chaotic and random, while english was projected as being modern, prestigious, and desirable. according to mitchell (1988), language was one of the most far-reaching strategies of the colonizers to change egyptian culture. he cites al-marsafi’s (1881) eight words, a book concerning the controversy of eight particularly powerful words that penetrated egyptian social and political life during british occupation: “nation”, “homeland”, “government”, “justice”, “oppression”, “politics”, “liberty”, and “education” (mitchell 1988:131). these terms were the new vocabulary of modern nationalism, and their perceived misunderstanding and misuse created a national crisis: the western values inherent in these eight words clashed with the cultural background of the egyptians. the impact of the colonizers on language and social identity was a breakdown of local culture (said 1975, 1979; abuhamida 1988; mitchell 1988). said’s (1975, 1979) controversial research on orientalism sharply criticizes western orientalists for perpetrating this linguistic imperialism and rendering the arabic language chaotic and random. said points to textual biases in literature about the orient. he claims, for instance, that arabs are metaphorically associated with hot-blooded sexual prowess (416), while institutionally or culturally “they are nil, or next to nil” (416). the use of sexual metaphor to describe male and female arabs was the orientalists’ way of “dealing with the great variety and potency of arab diversity, whose source is if not intellectual and social then sexual and biological” (416). these ideological forces gave rise to linguistic conflicts in post-colonial egypt: the desires for historical and linguistic nostalgia on one hand, and for modernization of language and society on the other. religious conservatives, fueled with anti-western sentiments and historical nostalgia, argue for a superiority of ca, and its purity is strongly anchored in muslim arab history, morality, and nationalism (suleiman 2003, 2004; haeri 2003). they relate “authentic” egyptian identity to islamic laws and values that are uncorrupted by the west (suleiman 2003, 2004; haeri 2003). this causes the religious conservatives to fight to keep ca undiluted with foreign borrowings. suleiman (2004) points to military warfare metaphors in their arabic rhetoric, such as “language regiment”, “defense of the national language”, and “enemies of islam” 2 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 3 (50), while they are projected as “holy warriors”, “garrisoned troops”, and “patrons” of classical arabic (49). any modernization of ca is lahn “linguistic corruption”, hadm wa-takhrib “sabotage”, or ghazw “invasion”, aimed at destroying the qur’an and the hadith “the prophetic traditions” (50). in contrast, pan-arab nationalists propose a united arabic language – msa to be the unifying force of all arabic-speaking people in the arab world. (abuhamida 1988; suleiman 2003; haeri 2003; amongst many). these idealistic nationalists argue for a written language that is mutually comprehensible in all arab nations and unifies the arab world. one proponent of this view is the egyptian government, which has been controlling the modernization of ca by overseeing institutions of learning, publishing, and social affairs since the middle of the nineteenth century (haeri 2003; van mol 2003). the government is mainly concerned with “revitalizing” ca as a means of achieving social, economic, and political progress for egypt. according to haeri (2003), state officials, egyptian intellectuals, educators, and high bureaucrats regard ca as too literary, flowery, and lacking in modern vocabulary needed for science and technology on a global scale. nevertheless, many voices point to the predicament that egyptian arabic, the vernacular, is ignored in writing and education, even though it is the mother tongue of egyptians and the lingua franca used in face-to-face interaction. in this view, egyptian arabic cannot be divorced from the identity of egyptian people and their local and national culture. linguistic regionalists, mainly arab and non-arab writers, are calling for the consolidation of spoken varieties at the expense of the standard variety (abuhamida 1988:42). according to abuhamida (1988), this call for linguistic regionalism coincides with political regionalism. lastly, proponents of a modern cultural and linguistic landscape in egypt advocate increasing use of english in many social domains in order to connect to the international community. their argument is that english has always been present in the postcolonial period. schaub (2000) states that “after a return to arabic and egyptian nationalism during the nasser period in the 1950s and 1960s, the situation again changed towards favoring english after the 1973 october war against israel” (228). he also explains that during the sadat years (1970-81), egyptian university students turned increasingly towards the united states. from 1974 on, “the u.s. agency for internal development has offered assistance to egypt in the training of public school teachers in english language instruction” (228). english in egypt is strategically used by the government and the media to achieve economic progress and to strengthen political and economic ties with the west. in sum, the interference of ‘modern’ western thought in colonial and postcolonial egypt acts as a powerful organizing force for present-day language ideologies: the devaluation of the local dialects as a result of both pan-arab nationalism and religious conservatism; an elevation of ca to a carrier of tradition and religious morals; the authority of msa as a contemporary standard variety able to reflect scientific and economic progress; and the use of english as 3 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 4 symbolic capital linking egypt to the “prosperity” of the west. proponents of each position use essentialized cultural differences between the east and the west as their logic and strategy to construct social action, morality, nationalism, and the “right” interpretation of language. in this highly contested, politicized language conflict over the “best” language variety, it is relevant to ask “which linguistic features are seized on, and through what semiotic processes they are interpreted as representing the collectivity” (woolard 1998:18). these topics are explored in section two, which uses linguistic and sociolinguistic research to provide a description of ca, msa, and egyptian arabic (ea) in the arabic diglossia in egypt and to show how language ideologies are linked to the linguistic structures of these language varieties. section three demonstrates how language ideologies are ranked, play a structuring role in every-day interaction, and shape a variety of communicative strategies. speakers use the shared historical, cultural, and linguistic background associated with varieties as an interactional resource in discourse. this dialectic aspect of ideology shows that “simply using language in particular ways is not what forms social groups, identities, or relations…; rather, ideological interpretations of such uses of language always mediate these effects” (woolard 1998:18). in such interactional uses speakers take advantage of the social, moral, and political attributes of each variety, which leaves a picture of dynamic reinterpretation of language use in egypt. the effects of these communicative strategies range from showing solidarity with the pan-arab nationalist ideology to transgressing social and geographic boundaries by tapping into western communicative styles. 2. arabic diglossia according to eisele (2002), dialect geography represented the most dominant form of linguistic analysis of arabic until the 1950’s (12). however, ferguson’s (1959) introduction of diglossia to the arabic sociolinguistic landscape “helped to crystallize modernist notions about this phenomenon and set the agenda for subsequent studies” (eisele 2002:12). ferguson (1959, 1996) conducted an extensive analysis of arabic diglossia. ferguson defines diglossia as a relatively stable language situation in which, in addition to the primary dialects of the language (which may include a standard or regional standards), there is a very divergent, highly codified (often grammatically more complex) superposed variety, the vehicle of a large and respected body of written literature, either of an earlier period or in another speech community, which is learned largely by formal education and is used for most written and formal spoken purposes but is not used by any sector of the community for ordinary conversation (ferguson 1996:34-35). ferguson claims that in arabic diglossia, ca is the divergent, highly codified, and superposed variety. it is seen as superior to vernaculars, such as ea, due to 4 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 5 widespread prejudices against vernaculars within the language community1. in line with ferguson (1996), this paper will use “h” for a high prestige language, such as ca or msa, and “l” for the low prestige dialects, such as ea. 2.1 h and l in diglossic situations, there is usually a belief that h is more beautiful, more logical, and more sophisticated (ferguson 1996; van mol 2003). this is true for arabic diglossia. according to haeri (2003), ca is often perceived as a “language whose aesthetic and musical qualities move its listeners, creating feelings of spirituality, nostalgia and community” (43). ca “socializes people into rituals of islam, affirms their identity as muslims and connects them to the realm of purity, morality, and god” (haeri 2003:43) and attributes of the language are often translated to the moral virtues of the user. ca also attained high prestige due to its rich literary tradition (ferguson 1996; van moll 2003, amongst many). there is a “sizable body of written literature which is held in high esteem by the speech community” (ferguson 1996:29-30). in contrast to ea, the orthography of ca is well established and “has a long tradition of grammatical study and a fixed norm for pronunciation, grammar and lexicon” (van mol 2003: 43). the fixed norms of ca were established as early as the ninth century, when ca was codified and, as the language of the qur’an, has been one of the major areas of study of muslims scholars ever since (parkinson 1991; van mol 2003). scholars have produced grammars, dictionaries, pronunciation manuals, and stylistic conventions that restrict variation and protect ca from the influence of modernity. this recalls woolard’s (1998:17) observation that “written form, lexical elaboration, rules for word formation, and historical derivation all may be seized on in diagnosing ‘real language’ and ranking the candidates” (17). somewhat contradictorily, haeri’s (2003) research also shows that many egyptians judge the spoken vernacular, ea, as the more beautiful variety, even for writing literature. she claims that “ordinary people describe it [ea] as easy, light, full of humor and more beautiful than other arabic dialects, as a habit and as the language of egyptians” (37). ea, or al-‘amiyyah “the common” is the mother tongue used for every-day communication in egypt and serves as a marker of egyptian identity and national culture (haeri 2003:37). nevertheless, ea is not generally recognized by religious scholars or pan-arab nationalists as a language of writing; since it has layers of lexical borrowings from coptic, turkish, persian, greek, italian, french, and english, it is criticized as “permissive”, “promiscuous”, or “weak” (38). furthermore, it is perceived as the 1 it is important to note, however, that not all egyptian dialects are valued equally. urban dialects, such as cairene arabic spoken in the city of cairo, usually have a higher prestige than rural dialects. 5 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 6 language of the ordinary egyptians on the streets, and the image of an ea speaker is that of a common or backward man (haeri 2003). this argument, however, is not sound since the extent of foreign borrowings into ca and their acceptability or naturalization is obscurred by the large timeframe in which it happened and “by the degree to which foreign elements have extended to be assimilated to the root-pattern system of the morphology and also by the organizing principles of arabic dictionaries, whereby assimilated borrowings are listed under a theoretical ‘root’” (305). for instance, the medieval borrowing di:ba:j “silk brocade” from the persian di:ba gave rise to the verb dabbaja “to embellish” (305). it then was reanalzed in the modern lexicon with the trilateral root d-b-j and took on arabic productive morphology, which can be seen in the derivation mudabbaja:t “figures of speech” (305). furthermore, concurrent with colonial times, “as western political, economic, and scientific ideas proliferated and ramified through the arab world, transliterated foreign words, especially in the sciences, and uncontrolled and sometimes inaccurate loan translations began to pour into written arabic” (308). this discrepancy between the ideologies and actual language features complicates the notion of the purity of ca. ferguson (1996) states that the communicative tensions between the h and l varieties in diglossia may be resolved by “the use of relatively uncodified, unstable, intermediate forms of the language” (31). msa is the most common intermediate form and is perceived as a “modern” version of ca. however, there are problems with defining msa, to which i turn to next. 2.2 msa many researchers show that there is no agreement as to what constitutes msa (parkinson 1991, 1992; haeri 2003; van mol 2003). opinions diverge as “to what extent and in what way modern arabic deviates from classical arabic” (van mol 2003:30) and to what extent colloquial or foreign linguistic elements, such as vocabulary, phonology, or syntax, are included. van mol (2003) claims that msa shows a large regional differentiation due to the influence of dialects and “it is not (yet) clear in what respect these regional varieties of msa exactly differ from each other and to what extent they differ as a group from their common origin, classical arabic” (4). this ambiguity has prompted many researchers to categorize fine-tuned intermediate language levels (van mol 2003, bassiouney 2006, blanc 1960, meiseles 1980, mitchell 1986, badawi 1973, amongst many). interestingly, parkinson’s (1991) approach to msa differs from these accounts in that he is not interested in categorizing levels but in how people perceive msa. he writes that “many of our problems in describing it [arabic] stem from the fact that it forms a relatively broad but indeterminate section of a much bigger continuum, and while there is general agreement about the continuum, there is little agreement about where the natural breaks are” (60). in 6 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 7 that sense, the value in parkinson’s (1991, 1991b, 1992) research lies in his claim that actual language use often is inconsistent with language ideology. in terms of ideology, he states that msa is an imperfectly known, but functional, part of most arabs’ communicative lives, associated with a rather high degree of linguistic insecurity, both respected and revered to the degree that it is viewed as a close relative or descendent of classical arabic, and despised and denigrated to the degree that it is taken to be a degeneration of classical arabic (1991b:48). speakers of msa see it as “a prescriptive form, a standard language that comes completely with a set of rules which define it” (32), but they differ as to what these norms are. it seems to be a moving target because it is an ideal variety that most people aim for in writing and speaking (51). however, modern fusha “recedes for some into a classicized, metaphor-laden, complex style not achievable by most modern writers” (51). consequently, shared perceptions of language features are shared manifestations of ideologies, which are not only derived from strictly linguistic categories but from “words and expressions as these are used by specific, historically located groups of users in the division of linguistic labor” (silverstein 1998: 128). shared ideologies, such as pan-arabism or linguistic purity, tie language to people, their culture, and their histories. there is a complex relationship between the structural features of a language variety and the ideologies associated with the variety, which is discussed in the final section of this paper. this section demonstrates silverstein (1998)’s claim that ideology “is defined only within a discourse of interpretation or construal of inherently dialectic indexical processes” (128). 3. language ideologies in interaction many researchers claim that in diglossia each language variety has a specific function (ferguson 1996). gumperz (1982) argues that “distinct varieties are employed in certain settings (such as home, school, work) that are associated with separate, bounded kinds of activities (public speaking, formal negotiations, special ceremonials, verbal games, etc.) or spoken with different categories of speakers (friends, family members, strangers, social inferiors, government officials, etc.)” (60). silverstein (1998) calls these distinct applications “default” functions that determine situationally appropriate language behavior. h, for instance, is often used in a “sermon in church or mosque, personal letter, speech in parliament, political speech, university lecture, news broadcast, newspaper editorial, news story, caption on [a] picture, caption on [a] political cartoon, and poetry” (ferguson 1996:28). indeed, studies show that in present-day egypt, msa is seen as the language of the media, education, and government, and is used in birth certificates, national identity cards, court trials, or deliberations in 7 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 8 the nagles el-shaab, egypt’s parliament (schaub 2000). homogeneity of language usage is also assumed for ca, which is reserved for the realm of religion. schaub (2000) suggests that since 88% of egypt’s population is muslim, a link between ca and religious practice is frequently reinforced. in contrast, l is used for “instructions to servants, waiters, workmen, clerks, conversation with family, friends, colleagues, radio ‘soap opera’, and folk literature” (ferguson 1996: 28). however, these distinct applications are not so straightforward in communicative interaction. speakers use language features from several language varieties in the same discourse in order to gain authority by tapping into specific language ideologies. according to gumperz (1982), this “reflects conventions created through networks of interpersonal relationships subject to change with changing power relationships and socio-ecological environments, so that sharing of basic conventions cannot be taken for granted”2 (95). violations of these conventions become a resource for effective communication. the following short case studies on text regulation (3.1), public speeches (3.2), and advertisement in the written media (3.3) show that, in arabic diglossia, the selective use of language features from different varieties signals as much information as the propositional content of the message: choosing features from one variety over another is a significant marker indexing the position of the speaker in society, their knowledge of political and religious values, or their aspiration for social mobility. 3.1 text regulation the egyptian government is the largest employer requiring msa in its public institutions, such as schools, and is actively involved in producing and controlling msa as a modern version of ca (haeri 1997: 800). its aim is to create a language that carries modern worldviews in such fields as science, politics, or arts, while at the same time to guard it from colloquial elements (haeri 1997). as woolard (1998) argues, “an ideology of ‘development’ is pervasive in postcolonial language planning, wherein deliberate intervention is deemed as necessary to make a linguistic variety suitable for modern functions” (21). haeri (2003) specifically points to state regulation of written texts through text editors, or professional “gatekeepers” (66), in egyptian publishing houses. their job is to change ca into 2 at this point, gumperz’s (1982) notions of situational and conversational codeswitching may be relevant. however, a discussion of codeswitching is beyond the scope for this paper. the immense complexity of integrating foreign or arabic semantic and phonological components into another code, or whether material switched must maintain all the characteristics of the original code have led to a variety of approaches (wilmsen 1996, van moll 2003, bassiouney 2006, amongst others). the treatment of codeswitching in arabic diglossia deserves another paper. 8 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 9 a medium for reporting all affairs of the world. specifically, they decide what constitute grammatical mistakes, infelicitous stylistic turns, inappropriate lexical choices, and suitable headlines, and they determine the form of the layout of the page (60). haeri highlights the importance of analyzing how the text became what it is, including “ideologies, institutional inculcations, and historical practices” (57). she describes various interpretations and changes in text against the background of the language ideologies associated with ca and ea. the cultural and political movement of text regulation is a central component of the modern arabic identity. her investigation of three text correctors reveals pan-arab nationalism (ilqowmiyya) as a crucial value (63). she also writes that “colonialism was also cited by all three as a reason for preserving and propagating classical arabic” (63). haeri’s account of reported speech is particularly revealing due to “the difficulties in accommodating, at one and the same time, incompatible, paradoxical and ambivalent ideologies with regard to both languages [ca and ea], and their relationships to culture…”(94). furthermore, she notes that the gatekeepers also have to consider the status and the personality of the person they are quoting. since ea has low prestige, “most prominent personalities cannot be represented as having spoken in that language” (98). to illustrate this point, haeri compares the representation in the press of speeches and interviews of four famous personalities: the egyptian president hosni mubarak, the novelist naguib mahfouz, the actor omar sharif, and the egyptian comedian adel imam. a short account of the reported speech of each interview is discussed next. in a televised meeting with writers and intellectuals in 1996, the president of egypt answered political and economic questions concerning egypt in informal ea (99). however, in the newspaper al-ahram, “every time the president was quoted, the quotation was in classical arabic” (99): president: da ihna rabbina bi-kul haazihi il-zuruuf il-sa‘ba… [ea] ‘why we thank our god (rabbina) that with all these difficult circumstances…’ al-ahram: wa qaala al-ra iis: wa nihmad allah ’anna(hu) bi-kul haadhihi al-zuruuf al-s‘aba… [ca] ‘and said the president: and we thank god (allah) that with all these difficult circumstances…’ (99) the demonstrative da “this” is used in ea idiomatically for emphasis, but is omitted in al-ahram’s quotation. furthermore, the ea pronoun ihna “we” is replaced by the ca nahnu “we”, and the vernacular rabbina “our god” is replaced by the ca allah ‘god’. haeri suggests that the translation from ea to ca “is meant for other arabs – serving the cause of pan-arabism – and the rest of the world” (100). 9 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 10 there are variations, however, in the ways in which reported speech is handled. when the writer naguib mahfuz gave a lengthy interview in ea to the magazine rose il-yussef, the bulk of mahfuz’s statements were printed in ca, with few ea interjections: naas rafdiin al-mugtama ‘people rejecting the society’ nisaq fii miin? ‘who should we trust?’ ehh ili biyihsal da? ‘what is happening?’ wa geeh min enn? ‘and where did it come from?’ wa issayy? ‘and how?’ haeri suggests that this is due to the fact that “the widely respected and famous novelist can be represented as having spoken a few brief phrases in his mother tongue” (101). the interview with the actor omar sharif was also conducted in ea, but here ea phrases are put in quotation marks. haeri states that “one reads the actor speaking in ca, and then a quotation mark appears with ea inside” (102). for instance, omar sharif: na‘am ana mizaaji jiddan… laew qumt min al-nowm wa sinna min asnaani tu‘limani “ab’a mish ‘aawiz’ashuuf hadd” ‘yes, i am very moody… if i woke up from sleep and one of my teeth were hurting “then i don’t want to see anyone”. haeri interviewed the text corrector who argued that ca is too formal and lacks feeling. “omar sharif would have ended up sounding more like a preacher or scholar [if he were quoted as speaking ca]” (103). in the last interview, the comedian adel imam also answered questions in ea and, in contrast to the others, the reported speech was printed entirely in ea. haeri claims that “it would probably be too much to represent the comedian as speaking anything other than his mother tongue, not only because that is how everyone knows him from movies and the media, but also because it seems to be judged as appropriate for a comedian not to speak ca” (104). in that sense, these four examples of reported speech show a definite hierarchy in which “the president of the country is not supposed to utter a word in the vernacular, the novelist a few more, the actor who has not lived in egypt for most of his life and is considered somewhat aloof and westernized can be represented as speaking both” (104). consequently, language regulation is not neutral or arbitrary, but “is based more often on political and social considerations than on linguistic or pedagogical factors” (woolard 1998:285). the ideologies associated with ca and ea determine “which linguistic features get selected for cultural attention and for social marking” (schieffelin and doucet 1998:285). determining factors 10 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 11 identified in haeri’s study are pan-arab nationalism and the ideologies of prestige or lack thereof associated with ca and ea, respectively, which compete and regulate language on a public and pervasive level. 3.2 public speeches public speeches in the arabic diglossia of egypt are generally delivered in msa or ca, but ea is “invading formal domains such as preaching, education, and all kinds of public speaking” (rabie 1991:422). however, the crossing-over of ea into other domains, such as public speeches, is not random and has significant consequences for the ideology of ea. while haeri’s research above painted a picture of ea being associated with low prestige and with persons of lower social status, holes (2004) shows that gamal abd al-nasir, the late president of egypt (1956-1965), used ea at mass rallies at the islamic al-azhar university in cairo to win sympathy and achieve his political goals. holes’ (2004) analysis of al-nasir’s speeches shows that often, in the same speech, “high-flown passages containing allusions to the glories of classical islam, or a peroration on the inevitability of socialist victory are leavened with others […] in which the rhythms and idioms of the street-wise cairene predominate” (holes 22). al-nasir’s speeches in pure ca usually entail that “mood and case endings are scrupulously respected throughout” (holes 2004:23) and “even phonological aspects... show almost no colloquial influence…”(24). al-nasir is using ca as “the language of political abstraction and symbol” (24), which allows him to present egypt as “a quasi-metaphysical essence,” defending its freedom and independence, or calling for peace, but declaring that it will fight (24). its people are never directly addressed, but presented as stylized “sons of egypt”, or the anonymous third-person “individuals” or “citizens” (24). using ca allows al-nasir to build up egypt’s moral strength by an impersonal revolution (24). however, in other speeches with the same content, al-nasir deliberately switches to ea to have a different communicative effect. he uses ea to indicate an audience-inclusive “we, the people”, a personalized picture of fighting “from house to house and village to village,” with the army “side by side with the people” (25). these forms are heavily ea in terms of syntax, word-level morphology, and suprasegmental features, which allows for a personal connection to the people. in the following excerpt, holes points to the role-relationship of the speaker and the audience, which indicates brotherly solidarity, expressed through firstand second-person pronouns in ea: ħanuħaarib “we will fight” iħna mustaʕiddiin, ʔayyuha l-ixwa ʔan nuqaatil “we are ready, brothers, to fight” 11 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 12 kuntu mawguud fi l-faluuga, zayyimaa intu taʕrufu “i was at faluga, as you know” ana mawguud maʕaaku hina fi l-qaahira “i’m with you here in cairo” ħanuqaatil, zayy maa ʔult ilkum imaariħ li aaxir nuʔtit dam “we will fight, like i told you yesterday, to the last drop of blood” (25) this performance provokes spontaneous audience participation. al-nasir’s aim is to inspire egyptians to pull together “in order to fight and defeat a foreign invader” (26). only ea seems to be able to convey this personal motivation. in that sense, al-nasir also uses ea to persuade the audience in the following way: 1. kaanu biyʔuulu nnu fiih ħurriyya siyaasiyya aw fihh dimuqraatiyya siyaasiyya [ea] “and they used to say that there was political freedom and there was political democracy” 2. wa laakin il-istiylaal wa l-/iqtaaʕ war ra/s il-maal al-mustayill qadaa ʕala kilmit id-dimuqraatiyya [msa] “but exploitation, feudalism and exploitative capital put an end to the idea of democracy” 3. illi /aaluuha [ea] “which they meant” 4. ʕalas&aan kida iħna bin/uul [ea] “so that’s why we say” 5. laa yumkin fi /ayyi ħaal /an yuqaal /anna hunaaka ħurriyya /illa /idaa tawaffarat ad-dimuqraatiyya s-siyaasiyya maʕa d-imuqraatiyya al/igtimaaʕiyya [msa] “it is impossible in any circumstances for it to be claimed that there is freedom unless political democracy exists alongside social democracy” (32). holes writes that the msa performances in (2) and (5) are presented as political axioms: democracy is indivisible, and cannot exist in an exploitative, feudalistic society (32). for sentence (1), on the other hand, holes notes that “the wrongheaded claim of an anonymous ‘them’ that ‘democracy already exists’, is delivered in rapid, conversational eca” (32). holes reasons that al-nasir’s reporting of the opposition’s claim in eca is “in itself a means of indirectly 12 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 13 signaling that what ‘they’ say is to be accorded less weight, and has less truth value than his axioms” (32). sentences (3) and (4) function as organizational: (3) anaphorically refers to “they” in sentence (1) while (4) faces both ways, linking what has gone before to what is to come (33). holes concludes that “these textually organizational elements are not in any way part of the message to be conveyed, but merely help the audience to recognize the real message – hence their rapid delivery in eca” (holes 2004:33). for personal, affective communication of domestic values, ʔaamiyya is used, which also organizes for the audience in “real time” the “timeless” fusha text (33). while msa has been associated in the psychology of the egyptian society as the language of “abstraction, idealization, and eternal values” (26-7), the switch to ea means that the relationship between speaker and audience moves from an impersonal one to one of friends. 3.3 media and advertising in contrast to ea’s influence in text regulation and public speeches, the media, advertising, and consumer culture are increasingly suffused with foreign borrowings, especially from english (van moll 2003, pimentel 2000, amongst many). suleiman (2004), for instance, discusses a shop sign, which appeared in heliopolis, a middle-class suburb of cairo, in the 1970’s: al-salam shopping centre li-l-muhajjabat “the peace shopping center for veiled women” (28). suleiman explains that al-salam is a popular term that arose in the 1970’s to refer to president sadat’s policy of pursuing a unilateral peace treaty with israel. the english term ‘shopping center’ is a reflection of “the strength of the westernoriented consumerism of the egyptian middle classes at the time, as well as the association between quality and foreignness” (28). the term al-muhajjabat “veiled woman” carries the ideologies of islamic traditionalisms in egypt due to the influence of the more conservative culture of the arabian peninsula and the belief in some circles that the muslim dress code “can eliminate the visibility of the socio-economic disparities between the rich and the poor in society” (28). van mol (2003) claims that foreign importations “are most easily integrated in the arabic dialects”3 (82). he states that editors of arabic language magazines and newspapers, especially those published in europe, are under heavy pressure to use european phraseology in their articles (83). schaub (2000) shows that the fashion bi-monthly cairo pose and the bi-monthly mother-to-be, which are entirely in english, are consciously targeted at egyptian readers (233). they are written for egyptian couples of the upper and upper-middle classes with disposable income. 3 furthermore, it is common that these foreign terms “gain currency in the spoken language before they find their way into writing so that they may be said to have come in not directly but via the colloquial” (76). 13 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 14 pimentel’s (2000) examined foreign and colloquial borrowings4 in newspaper display advertisement in contemporary egypt in daily newspapers such as al-ahrām, al-akhbār, al-gumhūriyya, and al-masā (37). he argues that egyptian advertisers strategically use english as a tool “to fulfill their commercial, informational, and ideological goals” (2). furthermore, instead of the fossilized, exclusively classical language sometimes described by western scholars, the written arabic of these ads proves flexible enough to address varying present-day needs: to convey new and innovative ideas, to communicate specifically to varying target audiences, and to be simplified as necessarily [sic] to facilitate communication (213). sometimes borrowings happen, according to wilmsen (1996), when “a term that is needed in a certain specialized use is borrowed from a code other than the baselevel colloquial – be it literary arabic or some second language” (84), such as when the members of a profession or other social group accept a foreign term as part of their jargon. along the same lines, pimentel (2000) claims that “borrowed lexical items in egyptian newspaper advertisements mostly refer to new technologies introduced from outside the arab world, having to do with computers, communications, electronics, and the automotive industry, as well as property, fashion, and entertainment” (56). such borrowings include intarnit “internet”, fīdiyū “video”, bārbīkiū “barbeque”, and sūbrānū “soprano” (56). in addition, borrowed trademarks can be generalized into expressing a foreign concept, as is the case with fiyūmāks “viewmax”, which is used generically for “television with vcr” (56). compound lexemes are also often borrowed, despite the fact that they are not common in arabic constructions without the possessive construction. these include, for instance, garāj sīl “garage sale” and stayshin wāgan “station wagon” (64). there are also compound loan structures mixed with arabic words, such as fagr shūbing santar “dawn shopping center”, markaz nū stār “the new star center”, tayyiba mūl “tayyiba mall”, and hadāyā al-krīsmās “the christmas gifts” (64). the frequency of borrowing english compounds is even more significant “given the availability of a common arabic structure to express similar semantic content” (68). according to pimentel (2000), these trends refer back to the “current 4 borrowing is usually treated as “the introduction of single words or short, frozen, idiomatic phrases from one variety into the other” (gumperz 1982:66). this usually happens when a term is needed in a certain specialized field and is borrowed from a code whose speakers are members of a profession or social groups that have it as part of their jargon. in the arabic diglossic environment, borrowings then mean “integration of a language l1 into another language l2, in which the borrowed item of l1 overtly takes on l2 characteristics, such as affixes and/or phonology” (wilmsen 1996:84). 14 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 15 trends toward globalization (al-cawlama) and privatization (al-takhaṣkhaṣa)” (212). as he claims, “globalization integrates egypt more and more fully into the single, global economy dominated by western capital and technology” (212). english in particular conveys an international feel, and some ideologies associated with commercial products are as important as the linguistic meaning potentially conveyed (213). this process, however, is not without ideological consequences because foreign borrowings, and more so colloquialisms, are conspicuous to some middle class egyptian readers. as pimentel’s (2000) research shows, “established borrowings apparently give little offense, while newly-borrowed items like tāym shīr ‘time share’ and garaj sīl ‘garage sale’ […] prompted scorn among conservative readers concerned with the welfare of standard arabic” (215). nevertheless, these terms are used in advertisements targeting upper-class audiences who perceive such usage as appropriately special and sophisticated. borrowings such as rīmūt kuntrūl “remote control” and sāntrāl lūk “central locking system” are likely the most effective here, conveying the sophistication appreciated by the upper-class (215). advertisements are “directed toward a population with specific cultural values within which the advertisements must function and also in the sense that those creating the advertisement operate within specific cultural norms” (11). in general, linguistic norms and sensibilities of the target audience are respected, but these high standards are sometimes maintained “at the expense of effective communication” (3). furthermore, some informants in pimentel’s study admit to the “usefulness of borrowings to describe technological innovations imported from outside the arab world” (211) and “the use of english as a symbol of modernity is more important than communicating through it” (211). the cultural message is more important than the linguistic one. as haeri (1997) argues, this revalorization is often widespread and reproduces the ideological superiority of the west. the relationship between the different languages and their ideologies is, in fact, a complicated one; advertising may well be tapping into popular, sometimes prescriptively incorrect, usage, but in doing so, “it reinforces and promotes such usage which may in turn become even more appealing to advertisers and may eventually change perceptions of linguistic norms” (pimentel 2000:3). furthermore, “a multi-level understanding of the registers embedded in written arabic begins to come into focus here, wherein the distinction between written and spoken arabic becomes less sharp and more fuzzy” (214). in terms of ideology, this may first have a certain promiscuous character to it, but eventually the status of the word is accepted and it assumes a stable form in the language (wilmsen 1996). it is important to note that despite both english and msa being h varieties, they have different symbolic capital, since only the upper-middle and upper classes have access to learning english in private schools as reported by haeri (1997). in this report, haeri offers a convincing argument against bourdieu’s (1977) claim that proficiency in the standard or national language is always associated with the highest symbolic capital and prosperity. the national 15 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 16 language in egypt, msa, does not imply as much prosperity as english does. as haeri writes, “if knowledge were always to equal power, the educated (lower) middle classes would have far more power than they do, since on the whole their proficiency in the official language is usually greater than the upper classes’” (haeri 1997:804). this situation causes resentment against english amongst the lower classes, and leads to wider resentment toward the government and foreign companies (schaub 2000:228). 4. concluding thoughts moment-by-moment language use often involves innovative and imaginative employment of language features. speakers in the arabic diglossia build on their own and their audiences’ understanding of language ideologies associated with each variety. cultural and linguistic frames have social histories, and this demands that we ask how seemingly essential and natural meanings of and about language are socially produced. although there are pre-determined domains in which ca, msa, ea, and english could be expected, i presented strong evidence that people tap into the communicative power of language ideologies in numerous ways. speakers not only use the language variety appropriate for a given situation, but they appropriate various language features for communicative effect. holes (2004) states that “although a person’s level of education, job, and social milieu will tend to determine the type and range of topics he/she usually talks about and who his or her regular interlocutors are, there is no automatic determination of style by the social identity of the speaker, except perhaps in the case of the completely illiterate” (15). i hope to have shown how language ideologies link language features to social processes in the arabic diglossia and how research focus on language ideologies could bridge linguistic and social theory. coupland (2001) charges that sociolinguistics has been largely unquestioning of the influences of modernity and globalization on language, ideology, and identity. in that sense, a focus on language ideology allows us to relate the microculture of communication and lived experiences to political considerations of power in a global world. references abuhamida, zakaria. 1988. “speech diversity and language unity: arabic as an integrating factor.” in ciacomo luciani and ghassan salam (eds.), the politics of arab integration, 25-42. london: croom helm. althusser, louis. 1971. lenin and philosophy and other essays. london: new left books. marsafi, husayn al-. 1881. risalat al-kalim al-thaman. cairo: mataba‘at al jumhur. badawi, a. 1973. mustawaya:t al-‘arabi:ya al-mu’a:sira fi: misr. cairo: da:r al-ma’a:rif. 16 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 17 bassiouney, reem. 2006. functions of code switching in egypt: evidence from monologues. leiden; boston: brill. bhabha, homi 1984. “of mimicry and man, the ambivalence of colonial discourse.” discipleship: a special issue on psychoanalysis, 28:125-133. blanc, h. 1960. “style variations in spoken arabic: a sample of inter-dialectal conversation.” in ferguson c. (ed.) contributions to arabic linguistics, 81-158. cambridge: harward univ. press. bourdieu, pierre. 1977. “the economics of linguistic exchanges”. social science information, 16.6: 645-68. bucholtz, mary. 2003. “sociolinguistic nostalgia and the authentication of identity.” journal of sociolinguistics 7(3): 398-416. bucholtz, mary and kira hall. 2004. “language and identity.” in alessandro duranti (ed.) a companion to linguistic anthropology. malden, ma: blackwell.  2005. “identity and interaction: a sociocultural linguistic approach.” discourse studies 7(4-5), 584-614. chakrabarty, dipesh. 2000. provincializing europe: postcolonial thought and historical difference. princeton, nj: princeton university press. coupland, nikolas. 2001. “introduction.” in coupland, srangi, and candlin (eds.) sociolinguistics and social theory, 1-26. harlow, uk: publishing house. eagleton, terry. 1991. ideology: an introduction. london: verso. eisele, john. 2002. “approaching diglossia: authorities, values, and representations.” in rouchdy, aleya (ed.) language contact and language conflict in arabic: variations on a sociolinguistic theme, 3 23. london: curzon. ferguson, charles. 1959. “diglossia”. word 15:325-40.  1996. sociolinguistic perspectives: papers on language in society, 1959 1994. new york: oxford university press. ferguson, james. 1999. expectations of modernity myths and meanings of urban life on the zambian copperbelt. berkeley: university of california press. friedrich, paul. 1989. “language, ideology, and political economy.” american ethnologist 91:295-312. foucault, michel. 1970. the order of things. new york: random house. foucault, michel, paul rabinow, and nikolas s. rose. 2003. the essential foucault. newyork: the new press. gumperz, john j. 1982. discourse strategies. cambridge [cambridgeshire]; new york: cambridge university press. haeri, niloofar. 1996. the sociolinguistic market of cairo. kegan paul international: new york.  1997. ‘the reproduction of symbolic capital’. current anthropology, 38(5): 795-817.  2003. sacred language, ordinary people. new york: palgrave 17 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 18 macmillan. haller, max. 2003. “europe and the arab-islamic world. a sociological perspective on the socio-cultural differences and mutual (mis)perceptions between two neighbouring cultural areas”. innovation: the european journal of social sciences 16(3): 227-252. hill, jane h. 1998. “ ‘today there is no respect’: nostalgia, ‘respect,’ and oppositional discourse in mexicano (nahuatl) language ideology”. in: schieffelin, bambi b., kathryn a woolard, paul kroskrity (eds.) language ideologies: practice and theory, 68-86. oxford university press. holes, clive. 2004. modern arabic: structures, functions, and varieties. washington, d.c.: georgetown university press. kandiyoti, deniz. 1996. gendering the middle east: emerging perspectives. syracuse, new york: syracuse university press. marx, karl and frederick engels. 1989. the german ideology. new york: international publishers. meiseles, g. 1980. “educated spoken arabic and the arabic language continuum”. archivum linguiticum 11(2): 118-148. mitchell, timothy. 1986. “what is educated spoken arabic?” ijsl 61: 7-32.  1991. colonizing egypt. cambridge: cambridge university press. nash, geoffrey. 1998. the arab writer in english: arab themes in a metropolitan language, 1908-1958. portland, ore.: sussex academic press. mol, mark van. 2003. variation in modern standard arabic in radio news broadcasts: a synchronic descriptive investigation into the use of complementary particles. leuven; dudley, mass.: peeters and departement oostere studies. parkinson, dilworth b. 1991a. “searching for modern fusha: real-life formal arabic.” al-arabiyya 24: 31-64.  1991b. perspectives on arabic linguistics: papers from the fifth annual symposium on arabic linguistics. amsterdam; philadelphia: john benjamins.  1992. “good arabic: ability and ideology in the egyptian arabic speech community.” ohak yonku/language research 28(2): 225-253. pratt, nicola. 2005. “identity, culture and democratization: the case of egypt”. new political science, 27(1): 69-86. pimentel, joseph j , jr. 2001. sociolinguistic reflections of privatization and globalization: the arabic of egyptian newspaper advertisements. phd. diss. university of michigan. rabie. medhat sidky. 1991. a sociolinguistic study of diglossia of egyptian radio arabic: an ethnographic approach. phd. diss. rampton, ben. 2001. “language crossing, cross talk, and cross-disciplinarily in sociolinguistics.” in coupland, srangi, and candlin (eds.) sociolinguistics and social theory, 261-296. harlow, uk. 18 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/4 doi: https://doi.org/10.25810/tdja-6p88 language ideologies in the arabic diglossia of egypt 19 rejwan, nissim. 1998. arabs face the modern world: religious, cultural, and political responses to the west. gainesville: university press of florida. said, edward. 1975. “shattered myths.” in naseer h. aruri (ed.) middle east crucible, 410-27. wilmette, il: medina university press.  1979. orientalism. vintage, publishers. schaub, m. 2000. “english in the arab republic of egypt”. world englishes 19(2): 225-238. schieffelin, bambi b. & rachelle doucet. 1998. “the ‘real’ haitian creole: ideology, metalinguistics, and orthographic choice”. in schieffelin, bambi b., kathryn a woolard and paul v kroskrity (eds.) language ideologies:practice and theory, 285-316. oxford university press. silverstein, michael. 1985. “language and the culture of gender: at the intersection of structure, usage and ideology.” in mertz, elizabeth and richard j. parmentier (eds.) semiotic mediation, 219-259. orlando, fla.: academic press.  1998. “the uses and utility of ideology: a commentary”. in schieffelin, bambi b., kathryn a woolard and paul v kroskrity (eds.) language ideologies: practice and theory, 123-147. oxford university press. suleiman, yasir. 2003. the arabic language and national identity. washington, dc: georgetown university press.  2004. a war of words: language and conflict in the middle east. cambridge university press. therborn, goran. 1980. the ideology of power and the power of ideology. london: nlb. thompson, john b. 1984. studies in the theory of ideology. cambridge, england: polity press. vološinov, v. n. 1973. marxism and the philosophy of language. cambridge, mass.: harvard university press. wilmsen, david. 1996. “codeswitching, code-mixing, and borrowing in the spoken arabic of a theatrical community in cairo”. in eid, mushira, & dilworth parkinson (eds.) perspectives on arabic linguistics ix: papers from the ninth annual symposium on arabic linguistics, 69-92. amsterdam: john benjamins. woolard, kathryn a. 1998. “introduction: language ideology as a field of inquiry”. in schieffelin, bambi b., kathryn a woolard and paul kroskrity (eds.) language ideologies: practice and theory, 285-316. oxford university press. 19 stadlbauer: language ideologies in the arabic diglossia of egypt published by cu scholar, 2010 colorado research in linguistics 6-2010 language ideologies in the arabic diglossia of egypt susanne stadlbauer recommended citation microsoft word cril_stadlbauer_comments_revised.doc microsoft word mcclelland-cril2021-proof_final.docx 1 a brief survey of anglicisms among spanish dialects jacob mcclelland university of colorado boulder in this survey we design a method for comparing the patterns of english lexical borrowing into dialects of spanish. using data collected from corpus del español actual (subirats & ortega, 2012), we compare the frequency of 28 anglicisms among dialectal subcorpora. we create a model that is normalized and standardized among the dialectal data. for each token or class of tokens a relative borrowing score (rbs) may be derived based on the deviation from the model baseline. overall rates of borrowing are thus modeled for each country, as well as semantic patterns of borrowing based on category. a comparison between the rates of particular anglicisms and native translation equivalents, is used to evaluate the semantic changes that accompany lexical borrowing. this rbs model is shown to concur with borrowing patterns attested in previous literature. it can serve as a broad tool for identifying areas for more detailed research. it may be adapted to more narrow investigations, such as the distribution of borrowings among sociolects. keywords: anglicism, lexical borrowing, dialectal variation, semantic modeling, spanish 1. introduction lexical borrowing is a significant process in the diachronic evolution of any languages which exist in prolonged contact. this borrowing is not always a symmetrical process and may be influenced by unbalanced power dynamics such as colonialism or the prominence of one language as a lingua franca for communicating broadly among speakers with diverse linguistic backgrounds. factors that impact the types and levels of lexical borrowing are prestige, domain specificity, and pragmatic constraints on cross-cultural or interlinguistic communication. the impacts of these factors are realized through a complex interplay of the introduction of terms through their idiosyncratic use by bilingual speakers and adoption and entrenchment of these terms into the monolingual lexicon (poplack et al. 1988). diachronic analysis of language contact may take on two methodologies, each relying on a different type of dataset. first, a researcher may examine longitudinal texts to evaluate the influence of language contact across extended time periods. second, one may examine latitudinal evidence of lexical borrowing across language dialects to make inferences based on the semantic and even grammatical impact that observed borrowing has had within a language. colorado research in linguistics, volume 25 (2021) 2 spanish and english coexist within many countries of the world. in 2015, there were 50 million heritage spanish speakers reported living in the united states, with 41 million reporting spanish to be their primary language spoken within the home. of these heritage speakers, 57% report to be fluent english speakers as well (us census, 2017). even for speakers non-fluent in english, exposure to the english lexicon is inevitable due to contact with bilingual and native english speakers. nearly 60% of the residents of spain report little to no proficiency in english (zafra, 2019). despite ranking 25th in english proficiency out of 35 european countries ( ef.edu, 2019), peninsular spanish shows high rates of borrowing in domains such as technology, tourism, and trade. the adoption of anglicisms into latin american spanish is by no means monolithic. the frequency of borrowings varies significantly by country, both in frequency and domain. crossdialectal borrowing, rather than direct contact with english, is a more prevalent source of anglicisms for some spanish dialects (mateescu, m., 2017). countries such as costa rica and panama are in contact with english speakers due to tourism and large numbers of american expatriates. panama has a particularly complex relationship with english due to the quasiimperialist influence of america related to the panama canal (alvarado de ricord, 1982). despite its long history of resistance to borrowing from english, trends in argentina now suggest that anglicisms are now becoming a marker of prestige in social stratification. (larsen, 2014). due to this complex diversity of influence that english has had in latin america, it becomes necessary to study anglicisms at the minimum of country level resolution. 2. project goals for this project, we use contemporary corpus data to perform an analysis of anglicisms as they occur in the united states, puerto rico, spain, as well as the 18 latin american countries where spanish is the majority language. we perform a quantitative analysis of the frequency of common anglicisms within text resources from these countries. we further analyze the frequency of these lexical borrowings as it relates to particular domains of use, such as food, technology, and clothing. we continue with a qualitative analysis on phenomena that we observe in the trends of this data. we then discuss particular word-pairs and what conclusions may be inferred from the variation observed among national dialects. a brief survey of anglicisms among spanish dialects 3 as to the part of our analysis that focuses on language contact, the expectation is that the frequency of anglicisms positively correlate to the english proficiency of speakers within a given country. we expect that the united states, puerto rico, and mexico will have the highest frequency of anglicisms, as there is significant social and geographical contact between english and spanish speakers in this region. we expect that nations that historically have had strained relationships with the united states, such as cuba and nicaragua, will exhibit a relatively lower frequency of lexical borrowing. this analysis will also examine the types of borrowings that are taking place and what effect they are having as to the semantic space of the different dialects. for terms with close semantic equivalence (synonyms/translation equivalents), we expect to see patterns that indicate what type of influence the adoption of the english lexical items is having on the historically spanish lexicon. we expect, as attested in previous literature (cabanillas et. al, 2007) (morin, 2006), to see a trend of spanish pulling anglicisms from the technical domain, which is dominated by global english. within countries with high levels of bilingualism and cross-cultural interaction, we expect to see evidence of anglicisms pushing their spanish origin equivalents into culturally specific domains. 3. methods we compile a list of 28 common anglicisms from contemporary blogs (kreisha, 2020), youtube videos (spanishpod101.com. 2020) and personal experience. these resources are all from united states sources and primarily reference anglicisms reported in united states and mexican-american spanish. for 16 of the anglicisms, we are able to identify semantically equivalent lexical terms of spanish origin. for the words bol and light, we select 2 semantic equivalents to compare a more complex semantic distribution. in total the wordlist contains the following number of anglicisms per category with the number of semantically close words of spanish origin shown in the parentheses: clothing 4 (1), technology 4 (1), social 8 (3), food 5 (5), general 7 (8). using the corpus del español actual (subirats & ortega, 2012) we compare the relative frequency of each lexical item in the list. this data is taken from the chart feature of the corpus interface and provides a frequency per million words for each of the 20 countries and the territory of puerto rico. we do an extensive statistical analysis on this data in order to identify trends and potential anomalies. for each term, we normalize to the average rate of occurrences across all colorado research in linguistics, volume 25 (2021) 4 dialects. it is not necessarily the case that this mean value represents any norm for borrowing per se, but it creates a useful scale for talking about dialects with relatively less (negative value) and more (positive value) frequency of borrowing. frequency patterns may vary significantly among lexical items do to the fact that each occupies a uniquely sized semantic space within the language and culture. to accommodate this feature, we standardize this space for each item by dividing each value by a standard deviation taken from the frequency variation for this term across dialects. for the purposes of this paper, we will refer to this value as a relative borrowing score (rbs). in principle this method is designed to evaluate any patterns of borrowing and other such variation across multiple dialects. after identifying terms that warrant a closer look, we perform a qualitative analysis using the sketchengine interface (kilgarriff et al, 2014). the spanish web 2018 (estenten18) corpus, available on sketchengine, is only divided into subcorpora for european spanish and spanish of the americas. sketchengine is useful for forming descriptions of the semantic range of terms, despite the limitations of its low geographical resolution. 4. results the data collected suggest substantive variation of english borrowings among spanish dialects. the mean frequencies of borrowings among dialects, overall and per category, were used as a baseline to evaluate the relative frequency of borrowing per semantic domain. an average semantic domain for each anglicism and spanish heritage word pair was created to establish a baseline to model semantic interaction. the overall composite frequency of these terms was considered as a baseline by which to evaluate semantic interactions among dialects. 4.1. quantitative analysis after the raw frequency of anglicisms are normalized and standardized, each country and category receive a composite relative borrowing score. this score reflects the number of standard deviations from the mean for frequency of anglicisms. the top scored countries are spain (0.67), chile (0.50), and peru (0.45). the lowest scored countries are nicaragua (-0.55), el salvador (0.59), and bolivia (-0.59). the united states (0.05) and mexico (-0.06) had an average number of anglicisms out of the countries studied. puerto rico (0.37) came in 6th in rbs, with a much higher rate of borrowing than the united states, to which it is a territory. a brief survey of anglicisms among spanish dialects 5 figure 1. relative anglicisms per dialect and category figure 2 shows the gradient of rbs per country and category, while figure 3 shows rbs per category in the top 10 borrowing countries (0.55 avg) compared to the bottom 11 borrowing countries (-0.32). the high frequency borrowers tended to adopt specialized terms categories such as clothing (0.52), food (0.49), and technology (0.40). the lowest frequency borrowers disfavored borrowing clothing (-0.43) and technology (-0.49) terms. this might suggest less cultural diffusion of global commercial products for these dialects. the lower gdp in these countries may also correlate with a lower demand for terms to signify technological products. heavy borrowing in technology is reflective of a pull pattern of borrowing, as the adopted words fill a semantic void. colorado research in linguistics, volume 25 (2021) 6 this was evidenced in the fact that we were not able to find translation equivalents for terms in this category figure 2. average anglicisms per country figure 3. high/low frequency borrowings per category the anglicism frequency scores compared to the education first english proficiency index (ef.edu, 201) appear to have a moderate positive correlation, but the size of this data sample does not establish a definitive level of significance (r=0.339, p=.156). there are a few notable exceptions to this trend: bolivia and paraguay have a negative borrowing score, but an above average proficiency in english. argentina has the highest proficiency score but has a moderate frequency of borrowings. it does hold that none of the top countries for anglicism frequency have extremely low english proficiency scores. a brief survey of anglicisms among spanish dialects 7 figure 4. relative borrowing score & ef english proficiency index scores in an analysis of anglicisms compared to their spanish origin semantic equivalents, we add the normalized frequencies of each token with the normalized frequency of its translation equivalent, thus estimating their combined share of each semantic domain in the dialect per dialect. a converse relative frequency would yield a zero score in this analysis. this may suggest that these terms are semantically equivalent; when the anglicism is adopted within a dialect, it replaces its translation equivalent within the semantic space. the bar for mail/correo (figure 5) reflects this balanced semantic distribution. when the combined relative frequency of these semantic pairs has a negative value, the terms occupy a smaller semantic domain. it is possible these terms belong to a less relevant semantic domain for that particular dialect. alternatively, it may indicate that there is some other term not reflected in this analysis, which compensates within this semantic space. this may indicate that there are lexical items missing from the analysis, which would increase the vocabulary within this semantic domain. the heavy left skew of carro/coche in figure 5 reflects this pattern. while the analysis suggests that these terms are equivalent in united states spanish, countries such as argentina that use the term auto show a significant negative skew. colorado research in linguistics, volume 25 (2021) 8 figure 5. semantic distribution of word pairs when this analysis yields a positive value, the terms occupy a larger semantic domain. this can be attributed to the terms not being an exact semantic match. the terms are near synonyms, but they are both used in the dialect therefore increasing the overall frequency. this can be seen in figure 5, by the strong right skew to piercing/perfocation. in several countries piercing is commonly used in the cosmetic sense, while perforcation is used to mean holes punctured in other contexts. this is an example of push borrowing, in that the native term is somewhat narrowed from its initial semantic domain. there are apparent trends within the data which show that our process for choosing semantic equivalents to be heavily skewed towards particular dialects. figure 6 demonstrates this with small and balanced bars for mexico and the united states, reflecting close matches in the word pairs. nicaragua and bolivia have large bars, indicating that the posited word pairs did not reflect the patterns of borrowing in these dialects. these are also the two countries with the lowest rate of anglicisms. a brief survey of anglicisms among spanish dialects 9 figure 6. semantic space of word pairs per dialect figure 6 also shows panama to be an interesting case. as shown by figures 1 and 2, panama has one of the highest rates of english borrowings. panama also has the broadest semantic space for the word pairs that we have considered. perhaps this reflects a phenomenon brought on through united states imperialism, where multiple registers of prestige inflate the semantic space. an analysis of two direct translation equivalent tokens is not always the most effective way to evaluate patterns of borrowing. terms among languages and dialects will have differently structured semantic space, as dictated by the culture in which they occur. to look at a more complex semantic distribution of equivalents within the domain of an particular anglicism, we consider bol/cuenco/tazón. figure 7 shows how the combination of all 3 terms best explains the interactions within the semantic space. this combination explains the makeup of this semantic space better than any of the proposed word pairs. colorado research in linguistics, volume 25 (2021) 10 figure 7. effectiveness of semantic modeling with three term split while this chart is a good indication of how we may analyze the semantic space it does not tell the whole story. there are two reasons that a dialect may vary in the frequency of terms for a particular semantic domain. if it is a cultural norm for food to be eaten by hand, then the semantic relevance of terms for utensils will be significantly diminished when compared to cultures that employ utensils. if a culture only has one word to describe utensils, then that word will have a large semantic domain as compared to a culture that distinguishes among spoons, forks, and knives. the two highest adopters of bol, spain and the united states, also have the highest combined rate for these terms. this suggests either that bowls have a larger semantic relevance in these cultures or that more items are considered to qualify within a larger semantic domain covered by these 3 terms, than would be indicated in other dialects. the united states is higher than average for the use of every one of these three terms. this suggests that these terms are distinctions of specific types that are relevant in the composition of this larger domain. latin american countries with high relative frequency of the term bol tend to have a strong negative correlation between its use and the combined value of the spanish origin terms. in bolivia, the highest latin american adopter of bol, the combination of the standardized frequency of these terms equals 0 as compared a brief survey of anglicisms among spanish dialects 11 to the mean. in chile, the second highest adopter, this same value is 0.13. this suggests that this term may be a direct semantic replacement in these dialects. 4.2. qualitative analysis for our initial analysis, cookie was placed in the category food. it quickly became apparent that the majority of the instances for the use of this term in fact belonged to the category of technology. while the singular form had a low rate of borrowing, cookies was the 6th most prevalent of the words analyzed. this demonstrates how words are not always borrowed from a language in their dominant sense. spanish term galleta is the translation equivalent of the english term cookie in its primary food sense. figure 8. spanish verb collocates of cookies (kilgarriff et al., 2014) in its technological sense, cookies refers to small files that are downloaded into an internet browser to track activity. spanish did not have its own word for these files so it borrowed the english term colorado research in linguistics, volume 25 (2021) 12 almost exclusively in the plural. speakers borrowed the term outright and did not reanalyze galletas in a technological sense. the term cookies has not significantly pushed into the semantic domain of galletas; it has instead been pulled from english to signify a new technological tool. figure 9: spanish collocates of ligero and light (kilgarriff et al., 2014) for the analysis of light we chose ligero and sano as translation equivalents. the attested sense of this borrowing was to describe diet products or meals that are not heavy. the choice of sano as an equivalent was not correct, in that spanish has a distinction sano/saludable that is comparable to the dwindling distinction in english healthy/healthful. the use of the term light in spanish is tied to commercial products from the united states. coca-cola and pepsi are 2 of the top nouns in spanish that are modified by this adjective. the term has undergone semantic creep in some dialects and may be used to refer to other healthful or low-fat items. some dialects, many with a high frequency of using light, have begun to reanalyze ligero to replace saludable, when referring to a healthful or a not heavy meal. in other dialects, the reanalyzed ligero may only be used to mean lightweight. 5. relevance & future work the broad nature of the analysis performed in the creation of a relative borrowing score is effective as a low-resolution signal of patterned borrowing as compared to a baseline. it was a brief survey of anglicisms among spanish dialects 13 effective in highlighting borrowing patterns already attested in previous literature (poplack et al. 1988). previously attested categories, such as clothing (balteiro, 2014) and technology (cabanillas et. al, 2007) (morin, 2006), were demonstrated to have high concentrations of borrowing among countries with high rbs. this method of analysis may be quite effective for investigating borrowing variation among sociolects. ostensibly, speakers from a particular region, due to shared environment and culture, would share a common semantic field within their lexicon. social factors can significantly affect the adoption of borrowed terms (sánchez, 2017). a shared regional dialect could create a more stable baseline for an rbs analysis, due to a reduced number of variables. any variation demonstrated through this analysis could provide clues as to specific correlates between social factors and lexical borrowing. in an analysis of borrowing patterns that reduces free variables, semantic modeling becomes more practical to describe variation. observations of how borrowings establish a position within a semantic field have increased pertinence. comparing the relative frequency of native and borrowed terms, as is done in this paper, may be paired with contemporary methods of vector space analysis (ganesh et al., 2017). 6. conclusion the analysis of the frequency of english borrowings into spanish dialects has yielded many informative results. it was surprising to find that the united states and mexico only shows a moderate rate of borrowing, when compared to the other countries analyzed. it may be necessary to examine the scope of the united states subcorpus; it may be the case that spanish material is catered to new immigrants with less exposure to english. bilingual speakers may be more likely to obtain news and other online media from english language sources. mexico has one of the lowest scores in the ef english proficiency index. this may explain why, despite close geographical proximity and cultural contact with the united states, the mexican dialect only has an average rate of english borrowings. cultural contact does seem to be a factor that influences the extent of english borrowings. cuba, bolivia, and paraguay have moderate to high rates of english proficiency, but due to limited cultural contact, they have a low rate of anglicisms within their dialects. colorado research in linguistics, volume 25 (2021) 14 as we further discuss in the conclusion, other factors may also influence the frequency of borrowings such as language attitudes. argentina has a history of laws discouraging language mixing. until recently, argentina even had regulations prohibiting non-spanish spellings for legal names (warren, 2015). recent adoption of anglicisms into the argentine dialect has been tied to a globalist trend in the higher economic classes. this history of language regulations and attitudes may explain why the argentinian dialect has an average frequency of anglicisms, despite having the highest english proficiency of spanish speaking countries. technology was not the most heavily borrowed category for countries with a high rbs. terms for clothing and food were borrowed more frequently. it is the case that the countries who borrowed technological terms the least, are also the ones that had the lowest rate of borrowings. these are also countries with low per capita gdp and low cultural contact with the united states. a promising addendum to our broader rbs model, is the method employed for analyzing relative semantic space. comparing the relative rates of borrowings and native terms appears to be a fruitful method for modeling semantic change due to language contact. while none of these models are definitive for identifying semantic shift, skews to the semantic distribution of the bars in figures 5 & 6 are suggestive as to whether replacement or semantic narrowing is taking place. the comparing the anglicism bol to two terms of spanish origin, was effective in demonstrating how complex semantic relationships can vary across dialects. finally, we observed how qualitative analysis may be necessary to describe certain patterns of borrowing. the categorical significance of a borrowing may be lost if an anglicism is analyzed based on its primary english sense. anglicisms may also have effects on spanish origin terms by causing a reanalysis of a spanish origin word to be used in an english sense. this paper functions as good proof of concept for this type of analysis. there are several tasks necessary to affirm the significance of these findings. for analyzing the relative rate of borrowings per country and per category, it will be necessary to develop a much broader list of anglicisms to be tested. the word list should take into account the most common anglicisms as researched for each dialect, not just mexico and the united states. it may be further necessary to evaluate each subcorpus to see if there is a balanced distribution of media type (news, blogs, government documents). for the semantic analysis, a more developed list of semantic equivalents is necessary. when a low skew is modeled for a word pair, additional words should be added for a more complete model. a brief survey of anglicisms among spanish dialects 15 for word pairs with a high skew, a qualitative analysis should be done to check for the semantic specialization that may be indicated. each dialect merits an evaluation of the factors that lead to its relative rate of borrowing. the reasons for variation posited in this paper were made through cursory observations based on the demographic information available for each country. economics, global relations, language attitudes, and relative contact exposure should also be accounted for. the most compelling application of this method is its ability to model the variation of borrowing among spanish dialects. this broad type of analysis accomplished through a relative borrowing score is effective for identifying borrowing patterns of interest for more detailed investigation. references alvarado de ricord, elsie. 1982 the impact of english in panama, word, 33:1-2, 97-107, doi: 10.1080/00437956.1982.11435725 balteiro, i., 2014. the influence of english on spanish fashion terminology:-ing forms. university of belgrade. faculty of economics. esp today journal of english for specific purposes at tertiary level. 2014, 2(2): 156-173. http://hdl.handle.net/10045/50263 cabanillas, isabel de la cruz redondo: cristina tejedor martínez; mercedes díez prados: esperanza cerdá redondo. 2007 english loanwords in spanish computer language, english for specific purposes, volume 26, issue 1, pages 52-78, https://doi.org/10.1016/j.esp.2005.06.002. ef.edu. 2019. ef english proficiency index a comprehensive ranking of countries by english skills. [online] available at: [accessed december 10th, 2020]. ganesh, barathi; anand kumar; soman, 2017. vector space model as cognitive space for text classification. arxiv preprint arxiv:1708.06068. mateescu, m. 2017. anglicisms in american spanish. special view on the media. contemporary readings in law and social justice, 9(2), 423-434. retrieved from https://colorado.idm.oclc.org/login?url=https://www-proquestcom.colorado.idm.oclc.org/scholarly-journals/anglicisms-american-spanish-special-viewon-media/docview/1973372050/se-2?accountid=14503 colorado research in linguistics, volume 25 (2021) 16 kilgarriff, adam; pavel rychlý; pavel smrž; david tugwell. 2014 itri-04-08 the sketch engine. information technology kreisha, meredith. 2020. “double agents: 68 sneaky english words undercover as spanish words,”. https://www.fluentu.com/blog/spanish/english-words-used-in-spanish/. larsen, jacqueline rae. 2014. social stratification of loanwords: a corpus-based approach to anglicisms in argentina. masters thesis, university of texas at austin. morin, regina. 2006. evidence in the spanish language press of linguistic borrowings of computer and internet-related terms. spanish in context, volume 3, number 2, 2006, pp. 161-179(19). john benjamins publishing company. https://doi.org/10.1075/sic.3.2.01mor muñoz-basols, javier & salazar, danica. 2016. cross-linguistic lexical influence between english and spanish. spanish in context. 13. 80-102. 10.1075/sic.13.1.04mun. poplack, shana & sankoff, david & miller, christopher. 1988. the social correlates and linguistic processes of lexical borrowing and assimilation. linguistics. 26. 47-104. 10.1515/ling.1988.26.1.47. sánchez fajardo, j.a., 2017. the anglicization of cuban spanish: lexico-semantic variations and patterns. universidad de alicante. departamento de filología inglesa. septentrio academic publishing. orealis: an international journal of hispanic linguistics. 2017, 6(2): 233-248. doi:10.7557/1.6.2.4120. http://hdl.handle.net/10045/71644 spanishpod101.com. 2020. english words used daily in spanish https://www.youtube.com/watch?v=ccurzcnm4ju subirats, carlos and marc ortega. 2012. corpus del español actual u.s. census bureau. 2017. “2017 american community survey 1-year estimates”. retrieved 2018-12-12 retrieved from https://archive.today/20200214011034/https://factfinder.census.gov/faces/tableservices/js f/pages/productview.xhtml?pid=acs_17_1yr_s1601&prodtype=table warren, sarah d. 2015. "naming regulations and indigenous rights in argentina." sociological forum 30, no. 3 : 764-86. accessed december 10, 2020. http://www.jstor.org/stable/43654132. a brief survey of anglicisms among spanish dialects 17 zafra, ignacio. 2019. “spain continues to have one of the worst levels of english in europe,” november 11, 2019. https://english.elpais.com/elpais/2019/11/08/inenglish/1573204575_231066.html. microsoft word quizar_sandoval-cril2021-proof_final.docx 1 the ch’orti’ project collaboration robin quizar rich sandoval metropolitan state university of denver dr. robin quizar and dr. rich sandoval are both alumni of cu boulder linguistics, and they are both affiliated with metropolitan state university of denver, robin as emeritus professor of linguistics and rich as assistant professor of anthropology. together they run a language documentation effort called the ch’orti’ project, of which robin is the director. robin worked extensively with the ch’orti’ (mayan) language and community in guatemala throughout the 1970s and 80s, helping to produce a number of language reference and revitalization materials. after retiring from msu denver, robin renewed this research, reconnected with the ch’orti’ community, and founded the ch’orti’ project in 2013 as a collaborative effort with msu denver’s ethnography lab. rich joined the project in 2017. given his background in linguistic anthropology, language documentation, and other relevant linguistics subfields, rich was a good fit to help robin run the project. the project’s accomplishments over the years are in large part due to the work of student assistants from the ethnography lab. one of the main goals of the project is to give these anthropology and linguistics undergraduate students real-world experience as well as the opportunity to develop a variety of practical and technological skills. the project has also involved other collaborators, including other scholar/researchers. because a primary focus of the ch’orti’ project is to support the ch’orti’ community’s own language revitalization efforts, including the reclamation of the classic mayan writing system, the project has undertaken a number of trips to the ch’orti’ communities of guatemala and honduras in order to learn about these efforts, conduct research, and otherwise develop community relationships. the essay elaborates on robin, rich, and other collaborators’ work with respect to these project activities and goals. it also provides background on ch’orti’ language revitalization efforts, general ch’orti’ language scholarship, and robin’s contributions to both. keywords: ch’orti’ mayan; language documentation, revitalization, reclamation; classic mayan; collaborative research 1. introduction although dr. robin quizar and dr. rich sandoval completed their graduate studies at different time periods, both are alumni of cu boulder linguistics. they currently work together on a longterm interdisciplinary community-oriented language documentation effort called the ch’orti’ project, of which robin is the director. the project is in association with metropolitan state university of denver (msu), where robin is emeritus professor of linguistics and rich is assistant professor of anthropology. in this brief essay, the authors describe the background, colorado research in linguistics, volume 25 (2021) 2 activities, and goals of the project, including information about the ch’orti’ language and community. in doing so, they hope to underscore how important it is for linguists working in such projects to collaborate and partner with community members and how beneficial it can be to involve a variety of colleagues and students. additionally, given the range of language-related phenomena involved in this particular project, the authors also hope to highlight the necessity for linguists to have – or be prepared to develop – skills across suband allied fields. in this way, they look back to the diverse opportunities and training that they were provided through their time at cu linguistics, a factor that should be evident in what follows. robin quizar received her phd from cu boulder, linguistics, in 1979, and returned to cu for an ma in anthropology, 1989. her doctoral dissertation, “comparative word order in mayan”, stemmed in large part from her work in guatemala with the proyecto lingüístico francisco marroquín (plfm). one of her tasks there was to teach linguistics to ch’orti’ speakers so that they could develop revitalization materials, such as a bilingual ch’orti’-spanish dictionary, a reference grammar, and translations of stories into their native language. in 1991 she was hired as a linguistics professor in the english department at msu denver. during her 20 years of teaching at msu, she headed the development of the linguistics program, which included two tracks for students to major in linguistics. upon retirement, she reconnected with the ch’orti’ community to support and assist their language documentation and revitalization efforts. she also continues to conduct related research, such as examining the relationship between ch’orti’ and other ch’olan languages and uncovering outside influences on ch’orti’ from xinkan, a non-mayan language (quizar 2020; quizar 2021). she is currently working on a historical reference grammar of ch’orti’. rich sandoval received his phd from cu boulder, linguistics, in 2016. his doctoral dissertation focused on the interactional and grammatical integration of spoken and signed language in arapaho (a plains algonquian language). this work was conducted in alignment with arapaho language documentation and revitalization efforts under the direction of dr. andrew cowell (cu linguistics) and with the support of northern arapaho community members. in 2019 rich was hired as a professor of anthropology at msu denver, making him the first and only fully dedicated linguistic anthropologist at msu. he is currently building up the linguistic anthropology curriculum with coursework that ranges from folklore to conversation analysis to epigraphy. the goal is to provide students with a broad foundation in linguistic analysis, including the ch’orti’ project collaboration 3 ethnographic, interactional, multimodal, critical, and historical approaches. his ongoing research and scholarship also cover various topics, including ch’orti’ and other language documentation work, multimodality in classic maya inscriptions, and historical sociolinguistic issues of spanishenglish contact in the american southwest. he is currently co-editing a volume of work on interactional approaches to language documentation (sandoval, williams, and sammons 2021). 2. the ch’orti’ project following her retirement, robin founded the ch’orti’ project in 2013 as part of msu anthropology’s ethnography lab and in cooperation with dr. rebecca forgash, the lab director. rich started working with robin on the project in 2017. in general, the ch’orti’ project is a collaborative effort involving a number of people and encompassing a variety of sub-projects and activities. a primary goal of the ch’orti’ project is to support the ch’orti’ community’s own language revitalization efforts, which also involves the reclamation of the classic mayan writing system and other aspects of their heritage. thus, one of the project’s focal activities has been yearly short trips to the ch’orti’ communities of jocotán, guatemala, and copán ruínas, honduras, to build and maintain relationships with the local people, as well as to conduct related research and learn how the project can be more supportive of their efforts. another of the primary goals of the ch’orti’ project is to provide undergraduate student research assistants with applied experiences relevant to their linguistics and anthropological coursework. student members of the project are hired as work study by the ethnography lab. depending on their interests and skills, they are assigned to work with others on different subprojects. some of this work involves linguistics research, while other work is supportive in other ways, such as data organization, website development, and transcribing. student assistants of the project also engage in presentations and other activities to learn about ch’orti’ language and culture, thereby becoming more invested in their work. of course, one of the best ways for students to become involved in the project is through travel to the ch’orti’ communities in guatemala and honduras, and many students have been able to do that throughout the years. as such, travel is a highlight, but the core of the student research assistant experience comes from year-round project work back at msu. under robin’s direction, ch’orti’ project work has included the preparation of ch’orti’ legacy texts, the development of pedagogical materials, and linguistics research involving the historical colorado research in linguistics, volume 25 (2021) 4 development of the language and its relation to other mayan languages, including its descendant relationship to classic mayan. she has also initiated and managed ch’orti’ project connections with the ch’orti’ community. currently she is working to establish project connections with the broader mayanist research community, which includes linguists, epigraphers, and anthropologists. rich has been working on developing ch’orti’ project initiatives involving language documentation and classic mayan writing. currently rich is heading the development of a ch’orti’ project website, which has the goal of making ch’orti’ language documentation, pedagogical materials, and related scholarship more accessible to the ch’orti’ community and interested scholars. dr. rebecca forgash (professor of anthropology, msu) has been intimately involved in ch’orti’ project work from the start. most recently this has involved ethnographic research on ch’orti’ understandings and language ideologies around speakership, sociocultural domains of use, and geographic distribution. jill scott (laboratory coordinator, department of sociology and anthropology, msu) supports student research assistants and helps administer project work. she is also part of the development team for the ch’orti’ project website. dr. andrew pantos (professor of linguistics, msu) directed and supported some past project research, notably involving the phonetic measurement of ch’orti’ vowel space. dr. marina gorlach (professor of linguistics, msu) has also participated in and supported project research. in what follows, ch’orti’ project activities will be elaborated on and contextualized within an overview of the ch’orti’ language, including revitalization efforts and related research. 3. ch’orti’ language and revitalization efforts the ch’orti’ (mayan) community is somewhat regionally isolated from the rest of the maya world, and the ch’orti’ language is currently spoken only in eastern guatemala near the honduran border. there are about 47,000 ethnic ch’orti’s living in guatemala and over 4,000 in honduras, but according to official census data, only about 15,000 ch’orti’s speak their native language, mostly in and around jocotán, guatemala. the number of speakers is unclear because the language is somewhat stigmatized, even within some of the ch’orti’ communities. notably, there is a long history of speakers not admitting to speaking the language given the possible risk of claiming an indigenous identity. the risk has been most severe during the waves of political and ethnic violence that characterize much of the past century in guatemala and honduras. the ch’orti’ project collaboration 5 for members of the ch’orti’ community, the current era of relative peace has meant an increased willingness to publicly identify as ch’orti’ and engage in associated activities. this factor is most notable in various community activist organizations that share a publicly stated goal of revitalizing and reclaiming the ch’orti’ language. however, although these organizations and associated initiatives have had some successes, they face a variety of difficulties that are best understood through a survey of their development. initial work at language revitalization started in the 1970s with the proyecto lingüístico francisco marroquín (plfm), an organization based in the guatemalan cities of huehuetenango and antigua. the plfm supported the revitalization efforts of numerous mayan languages by hiring linguists to work on languages and train community members in a way that would enable them to develop their own language-learning pedagogical materials. robin worked with the plfm and a team of ch’orti’ speakers during 19751979, which aided in the development of a plfm ch’orti’ dictionary (pérez et al. 1996), a reference grammar (pérez 1994), and a set of stories in ch’orti’ (pérez 1996). the idea was for the plfm-trained ch’orti’ linguists to use their knowledge to teach ch’orti’ linguistics to other members of the ch’orti’ community. the early success of plfm in training local ch’orti’ speakerlinguists is evidenced by continued trainee involvement in jocotán’s bilingual education programs, as well as adult literacy programs, even after plfm had become less influential in revitalization efforts. since the signing of the 1996 peace accords in guatemala, the government has underwritten bilingual education programs in state schools. additionally, the government pledged support to the academia de lenguas mayas de guatemala (almg), a community activist organization formed in the late 1980’s with the mission to support the revitalization and maintenance of the country’s indigenous languages, in alignment with the pan-maya movement. with a nod to the plfm, the almg worked to train native linguists and publish pedagogical language materials. most significantly, the almg used the plfm’s original orthographic conventions to develop a standard alphabet for all of guatemala’s mayan languages. this move was quite important for ch’orti’ revitalization efforts, as the various inconsistent orthographies represented in teaching materials presented unnecessary difficulties for learners (and teachers). however, the guatemalan government no longer provides the same level of financial support to the almg, and so the ch’orti’ branch of the almg in jocotán lacks the funds needed to further develop much-needed teaching materials. the ch’orti’ almg continues to do what they can to colorado research in linguistics, volume 25 (2021) 6 support teacher training, both for public education teachers as well as for those involved in more community-based programs. one implication of this situation for the community is that the government’s primary goal is to use bilingual education as a pedagogical tool to increase children's abilities in standard spanish as opposed to helping to revitalize and maintain the ch’orti’ language. this is most evident in that ch’orti’ language teachers, although being ethnically ch’orti’, often cannot speak the language themselves. despite the existence of ch’orti’-speaking administrators trained in linguistics, guatemalan governing institutions and other aspects of the sociopolitical system unfortunately make it extremely difficult for actual fluent ch’orti’ speakers to gain the credentials to teach ch’orti’. because schools lack resources to otherwise support teachers in any efforts to develop ch’orti’ fluency among students, those who are credentialed to be ch’orti’ language teachers mostly focus on basic tasks that don’t require fluency, notably teaching the letter-sound correspondences of the ch’orti’ alphabet. the result is that ch’orti’ language lessons are less about language revitalization and more about functional literacy, skills that are transferable to spanish or any other language with a romanized writing system. another hurdle facing the ch’orti’ almg and other revitalization efforts is that, for most ch’orti’ adults, the traditional language only has nostalgic value, in contrast to the social, political, and economic value associated with spanish. thus, adults often insist on speaking only spanish with their children to better prepare them for success in school. without real government support and without a common sense of urgency for language revitalization, the almg and other ch’orti’ community activist organizations are tasked with not only ch’orti’ language education but also with building a stronger sense of community heritage and identity around the language. 4. reclaiming ch’orti’ heritage and classic mayan hieroglyphs in support of these goals, one emerging area of interest for ch’orti’ activists has involved the reclamation of pre-columbian classic mayan writing and other traditions. relative to other areas of the maya world, the ch’orti’ interest here is recent, owing in large part to a contemporary body of developing research showing that ch’orti’ is the living mayan language with the closest descendent connections to the written language of the classic maya civilization, c. 250-900 ce. mayan linguists and other mayan scholars have claimed for decades that classic mayan writing represents a language from the ch’olan branch of the mayan language family. however, controversy exists regarding which modern ch’olan language has the most direct claim to the the ch’orti’ project collaboration 7 language of the hieroglyphs. some mayan scholars posit that the ch’olan branch was a single language, i.e., proto-ch’olan, during the classic period (e.g. mora-marín 2009). linguistic variation in the corpus of inscriptions would thus be due to dialectal variation across the classic maya area. under this view, all ch’olan languages share equally in the glyphic heritage. a competing and increasingly more dominant hypothesis holds that the ch’olan languages had already split up into eastern and western language groups by the time of the classic maya, the writing system being based on eastern ch’olan, which includes ch’orti’ and the now-extinct ch’olti’. robin is involved in working to understand the relationship between these two eastern ch’olan languages. some linguists claim that ch’orti’ is best described as a descendant of ch’olti’, the two languages thus representing different historical stages (robertson 1998; houston et al. 2000; robertson and law 2009). others support a more traditional view that the two were distinct sister languages. current research by robin leads her to support the latter view of separate languages. her findings indicate that in many respects ch’orti’ is more conservative than ch’olti’ (quizar 2020) and that certain grammatical features in ch’orti’ were likely the result of contact with speakers of the non-mayan xinkan language (quizar 2021). regardless, because ch’orti’ is both a ch’olan language and an eastern ch’olan language, the consensus view now is that ch’orti’ represents a descendant language of classic mayan. informed by this research and its historical perspective, there is a growing sense among the ch’orti’ that they have more claim than other mayan communities to mayan glyphic writing and other associated classic traditions as part of how they construct ch’orti’ identity. here, too, community activists have a steep road to climb. k’iche and other non-ch’olan mayan communities have been involved in the reclamation of mayan glyphic writing for several years and have popularized their own association with the classic maya, relying on this claim for tourist dollars. moreover, the ch’orti’ language (and other ch’olan languages) is erased from narratives of the classic mayan inscriptions in many museums, including those in the famous classic maya site of copán, located in ch’orti’ territory. additionally, there are teachers with well-developed pedagogical resources throughout the maya world who are engaged in supporting community efforts to reclaim classic mayan glyphic writing, but the location of the ch’orti’ community means they once again suffer due to their geographic separation from the rest of the maya world. regardless, ch’orti’ activists understand that developing this connection with their classic maya heritage will help to create a stronger sense of ch’orti’ community. given the situation in colorado research in linguistics, volume 25 (2021) 8 other maya communities, ch’orti’ activists also understand that such reclamation efforts could do well to increase the community visibility with respect to tourists and researchers, giving the ch’orti’ language economic value as well. the desire and move to claim and teach mayan glyphic writing and other aspects of classic maya traditions are therefore seen by many as integral to broader language revitalization and community development efforts. 5. language documentation, other project work, and looking forward given this social and historical context involving ch’orti’ language revitalization, participants in the ch’orti’ project are constantly positioning themselves to understand how their work can directly or indirectly support community efforts and resultant needs. specifically, a focal part of ch’orti’ project travel has been to visit and work with the ch’orti’ almg in jocotán, guatemala. for example, the almg has hosted a ch’orti’ project workshop discussing the use of ch’orti’ texts as pedagogical materials in the schools and another one introducing classic mayan glyphic writing to the local community. plans are in the works for future workshops on mayan glyphs and other language-oriented topics of community interest and need. much of the work done by robin and other ch’orti’ project members during these visits, however, has been to assess the needs of the almg and other community-based organizations. for example, some of the earliest work of the ch’orti’ project involved finding out what kind of language pedagogical materials were needed during one trip, developing them back in denver, and then, for the next trip a year later, delivering those materials. currently, rich is leading a ch’orti’ project team in developing a ch’orti’-language website, one of the main goals being to provide support to the almg and other ch’orti’ language initiatives. the plan is for the website to provide pedagogical materials as well as reference information on ch’orti’ linguistics and classic mayan writing, among other things. the ch’orti’ project also supports ch’orti’ language revitalization efforts by helping to develop the body of ch’orti’ language documentation and linguistics research. very early on, spanish priests engaged in ch’orti’ linguistic scholarship as part of their goals of spreading catholic doctrine amongst natives. for example, morán in the 17th century worked with ch’olti’ (closely related to ch’orti’), documenting vocabulary and texts and describing much of its grammar. in the early 20th century, the ethnographer charles wisdom recorded ch’orti’ texts and took detailed notes on the language, although this work was not published (but it has always been the ch’orti’ project collaboration 9 easily accessible to researchers). a more significant body of research on ch’orti’ language was developed by john fought, working in the jocotán area in the 1960’s. his work includes a reference grammar and a varied and sizable corpus consisting of dozens of texts on topics that include religion, home life, and agriculture (1972). starting in the 1970’s, john lubeck, as a summer institute of linguistics initiative, settled in jocotán to create and disseminate a ch’orti’ translation of the bible. he has also published a pedagogical grammar with diane cowie (1989). since then, a variety of linguists and anthropologists, including otto schumann, kerry hull, cédric becquey, and robin quizar, have recorded ch’orti’ texts, compiled wordlists, and engaged in other language documentation research, most of which is not published, at least not yet. kerry hull, however, has recently published an extensive ch’orti’ dictionary (2016). the ch’orti’ project has been working in a number of ways to build on this body of work. its biggest endeavor to date has involved robin and student research assistants in the preparation of legacy ch’orti’ texts. this work includes the digitization and transliteration into modern ch’orti’ alphabet of 40 texts on a variety of topics, originally collected and phonetically transcribed by john fought (1972). other work includes a similar preparation of texts about daily life and culture, which were collected by charles wisdom (1950). wisdom’s original corpus is comprised of 180 handwritten pages, including english glosses for each word. the ch’orti’ project has also prepared texts collected by vitalino pérez (1996). these legacy texts are now much more useful for scholars and accessible to the ch’orti’ community. during the project’s 2014 trip to guatemala, copies of these prepared texts were presented and given to the almg. the ch’orti’ project website will make the texts even more accessible. under rich’s guidance, the ch’orti’ project has also taken some initial steps to develop a ch’orti’ language database consisting of video-recorded everyday interactions. in 2018, rich and robin conducted meetings with a variety of ch’orti’ community members to gauge interest in such a documentation project. the response was overwhelmingly positive and encouraging, especially given the potential for education of having ready examples of fluent speakers using everyday ch’orti’ with one another. members of the ch’orti’ almg partnered in the endeavor, making plans to help the ch’orti’ project start video recording during the next year’s travel. unfortunately, a tragedy in the community as well as us government travel advisories temporarily halted a continuation of the documentation effort in 2019. the covid-19 pandemic has meant that it may be a while before it’s feasible to restart the effort. colorado research in linguistics, volume 25 (2021) 10 while this setback has been a bit unsettling for the ch’orti’ project, such complications are not unexpected. as with any other dynamic community-engaging linguistics project, robin, rich, and other ch’orti’ project members expect any given task to present unforeseen challenges, even tasks with a seemingly straight-forward process. thus, as much as project members are engaged in actual tasks, they are also soliciting feedback on works in progress, experimenting with and developing methods and means for achieving goals, and brainstorming ideas for future efforts. current project brainstorming involves expanding internet presence beyond the website so that (potential) ch’orti’ speakers can engage in other ways with the language and do so more conveniently. ideas include online language-learning games and translation programs. of course, such tasks will first require ch’orti’ project members to gauge ch’orti’ community interest and concerns as well as overall efficacy of such efforts. another, potential future endeavor involves engaging with the various museums in guatemala, honduras, and the us where ch’orti’ language and culture are not accurately or fairly represented. finally, project members are considering ways that they might engage with and support ch’orti’ migrant communities in the us. regardless of when, if, and how these ideas are realized, the ch’orti’ project is set to continue making progress by working to develop ch’orti’ scholarship and support the revitalization efforts of the ch’orti’ community. references fought, john g. 1972. chortí (mayan) texts (i). philadelphia: university of pennsylvania press. houston, stephen, john s. robertson, and david stuart. 2000. the language of the classic maya inscriptions. current anthropology 41(3):321-55. hull, kerry. 2016. a dictionary of ch’orti’ mayan-spanish-english. salt lake: university of utah press. lubeck, john and diane cowie. 1989. método moderno para aprender el idioma chorti’. guatemala: instituto lingüístico de verano. mora-marín, david f. 2009. a test and falsification of the ‘classic ch’olti’an hypothesis’: a study of three proto-ch’olan markers. international journal of american linguistics 75(2):115-157. pérez martínez, v. 1994. gramática del idioma ch’orti’. antigua, guatemala: proyecto lingüístico francisco marroquín. the ch’orti’ project collaboration 11 pérez martínez, vitalino. 1996. leyenda maya ch’orti’. guatemala: proyecto lingüístico francisco marroquín. pérez martínez, vitalino, federico garcía, felipe martínez, and jeremías lópez with robin quizar (linguistic advisor). 1996. diccionario del idioma ch’orti’ (jocotanchiquimula, ch’orti’-español). antigua, guatemala: proyecto lingüístico francisco marroquín. quizar, robin. 1979. comparative word order in mayan. phd dissertation. university of colorado at boulder. quizar, robin. 2020. tracing the ch’orti’ antipassive system: a comparative/historical view. international journal of american linguistics 86 (2): 237-283. quizar, robin. 2021. early split between ch’orti’ and ch’olti’: xinkan influence on ch’orti’ verbs. manuscript submitted for publication. robertson, john s. 1998. a ch’olti’an explanation for ch’orti’an grammar: a postlude to the language of the classic maya. mayab 11. unam: sociedad española de estudios mayas. robertson, john s., and danny law. 2009. valency to aspect in the ch’olan-tzeltalan family of mayan. international journal of american linguistics 75(3):293–316. sandoval, rich. 2016. gesture-speech bimodalism in arapaho grammar: an interactional approach. phd dissertation. university of colorado at boulder. sandoval, rich, nick williams, and olivia sammons (editors). 2021. interactional approaches to language documentation. a language documentation & conservation special publication. in preparation. wisdom, charles. 1950. materials on the chorti language. university of chicago microfilm collection. the semantic contribution of complementizers and complementation type: the case of bolanci na 1 awad: the semantic contribution of complementizers and complementation type published by cu scholar, 1995 2 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/2 doi: https://doi.org/10.25810/n62j-v827 3 awad: the semantic contribution of complementizers and complementation type published by cu scholar, 1995 4 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/2 doi: https://doi.org/10.25810/n62j-v827 5 awad: the semantic contribution of complementizers and complementation type published by cu scholar, 1995 6 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/2 doi: https://doi.org/10.25810/n62j-v827 7 awad: the semantic contribution of complementizers and complementation type published by cu scholar, 1995 8 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/2 doi: https://doi.org/10.25810/n62j-v827 9 awad: the semantic contribution of complementizers and complementation type published by cu scholar, 1995 10 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/2 doi: https://doi.org/10.25810/n62j-v827 11 awad: the semantic contribution of complementizers and complementation type published by cu scholar, 1995 12 colorado research in linguistics, vol. 14 [1995] https://scholar.colorado.edu/cril/vol14/iss1/2 doi: https://doi.org/10.25810/n62j-v827 colorado research in linguistics 1995 the semantic contribution of complementizers and complementation type: the case of bolanci na maher awad recommended citation tmp.1537644141.pdf.sirop microsoft word adams-cril2021-proof_final.docx 1 a pilot study on voice-conditioned vowel raising comparison in the hockey community of practice sarah adams university of colorado boulder this project1 investigates vowel raising in the hockey community of practice, specifically englishspeaking athletes who are were born and raised in central canada or in the northern cities region of the united states. research done on “canadian raising” in north america looks at the type and location of use; i chose a narrow group of speakers based on the dominance of canadian-born athletes in the community and the stereotype that they “sound canadian.” the effect of the voicing of a post-vocalic consonant could have a deeper sociolinguistic connection regarding identity and speakers’ occupation. i compared their place of origin with their current place of residence to determine who has more contact with canadian english speakers and would be influenced by their speech patterns to focus on the sociophonetic question. my research question is about the cross-linguistic influence of english-speaking hockey players who are, through professional contact, exposed primarily to canadian monolingual english speakers. the present study is motivated by sociophonetic considerations, and a focus on groups of athletes for sociolinguistic reasons influenced my choice to investigate this question. because the raising can also be found in inland north american english, i chose to include athletes from that region in this data. i recorded and measured examples of the athletes’ connected speech from video interviews and ran t-tests comparing the canadian-born and american-born athletes together, and then dividing them by current place of residence for additional comparison. the results show promising preliminary results in line with speakers residing in canada “sounding more canadian” regardless of place of origin, however the t-tests are not significant. more data being collected in an appropriate setting would enhance these results and further study of multilingual athletes would further inform the sociophonetic aspect pertaining to sense of belonging and conformity in the hockey community of practice. keywords: sociolinguistics, canadian english, sports linguistics, sociophonetics, phonetics 1. introduction vowel raising is an observable phenomenon where vowel sounds are changed due to the tongue being higher in the mouth. in north american english dialects, this can be observed across parts of canada and the northern united states, even though it is most commonly called “canadian raising” (chambers 1973). this project investigates the effect of the voicing of a post-vocalic consonant in conjunction with the place of origin and current place of residence of several north colorado research in linguistics, volume 25 (2021) 2 american athletes. i choose on the hockey community of practice because of the dominance of canadian-born athletes in the sport, as well as the stereotype of canadian speech associated with the community, regardless of actual place of origin. there is also a deeper sociolinguistic question about identity related to my speakers’ occupation. if they are “expected” to sound canadian, will that contribute to more evidence of raising? or, if they feel connected to their place of origin, would that supercede the desire to fit in with a canadian-sounding majority? do canadian-born players living in some parts of the united states also desire to fit in there and to not exhibit raising? my research question is about the cross-linguistic influence of english-speaking hockey players who are, through professional contact, exposed primarily to canadian monolingual english speakers. the present study is motivated by sociophonetic considerations, and a focus on groups of athletes for sociolinguistic reasons influenced my choice to investigate this question. research by dailey-o’cain (1997) made the observation that while [aɪ] raising is part of the dialect in the northern united states, the [aʊ] raising can be found there as well. because the raising can also be found in inland north american english, i chose to include athletes from that region in this data to see if they exhibit influence of canadian raising due to their community of practice. when i was determining which athletes to find speech tokens for, i organized a list based on whether they are originally from central canada or the northern cities geographic region, and then additionally whether they play for a team in canada or the northern cities geographic region. i labeled teams outside of the canada/northern cities region as other. i narrowed the list to players who have played either exclusively or primarily for their current team and who, to the best of my knowledge, had spent the majority of their formative years in their place of origin to reduce the amount of variation. i focus on the two sound variations noted in canadian raising: [aɪ] and [aʊ] before voiceless consonants, and i also measure these before voiced consonants as a means of comparison. if my hypotheses are accepted in the data, that would indicate that this vowel raising is found in a community of practice which gives athletes in the canada/the northern cities more exposure to and reinforces their use of “canadian raising.” my hypotheses are as follows: • hypothesis 1: canadians will exhibit more [aɪ] and [aʊ] raising before voiceless consonants than americans. a pilot study on voice-conditioned vowel raising comparison 3 • hypothesis 2: players on canadian/northern cities teams will exhibit more [aɪ] and [aʊ] raising before voiceless consonants than other teams; a. canadian players on canadian/northern cities teams will exhibit more raising than canadians on other teams. b. american players on canadian/northern cities teams will exhibit more raising than americans on other teams. • hypothesis 3: canadians will show a smaller difference based on team location than americans. this paper consists of an explanation of my data collection methods, more specific information about players, and a list of data tokens in section 2. in section 3 (results and analysis), i include simple tables to illustrate the measured vowels in direct comparisons, and i conclude in section 4 with a general discussion of my results, including limitations and additional questions for future research. 2. methods the experimental materials included the words listed below, organized by diphthong and voiced or voiceless. i was unable to retrieve recordings of every word for every player, so while i was able to measure “about” [aʊ] raising before voiceless consonant [t] for every player, i was not able to find the same words for each environment and will be generalizing across phonological contexts. colorado research in linguistics, volume 25 (2021) 4 table 1. target word list [aɪ] before voiceless excite(-ing, -ed) hype right guys ice nice quite alright night/s bite/s type fight might mic light advice biased [aʊ] before voiceless about out scouting house throughout without [aɪ] before voiced bide ride tied side stride describe inside divide idea pride tried wide outside driveway realize wives tiger [aʊ] before voiced proud thousand loudest the variables that are being manipulated can be seen in my choice of subjects and the recorded speech. i chose to analyze the speech of canadians from the central provinces of canada and americans from the northern cities region, minnesota and wisconsin specifically. i chose athletes from these regions because of the assumption that they will exhibit more [aɪ] and [aʊ] raising before voiceless consonants than athletes from other regions so i will be comparing this group expecting to see raising in all speech examples. however, i am hypothesizing that the athletes who play on teams in other regions of the united states like new york, philadelphia, dallas, and denver will all exhibit less raising as a result of this contact. therefore, i am choosing words from their connected speech which contain the diphthongs in question before voiced [b], [d], [g], and [z] sounds and voiceless [p], [t], [k], and [s] sounds. the measurement criteria changed only slightly during the data collection process. initially, i was capturing the connected speech around the diphthong, and then cutting out the extraneous words in the recording to only get the isolated word for measurement. however, i realized that this was taking a significant amount of time and it wouldn’t actually impact the results. so i a pilot study on voice-conditioned vowel raising comparison 5 transitioned to simply measuring the diphthong from its original recording. when measuring the diphthong, i looked to the nucleus for the raising, so i needed to measure in the first third and the second third of the whole vowel sound. i measured these consistently and recorded the formant measurements in praat and labeled the nuclei for testing. 3. results in this section, i consider the three hypotheses found in section 1, labeled 3.1, 3.2 with subsections 3.2a and 3.2b, and 3.3 and i used t-tests to analyze the data. i included graphs to illustrate my findings. 3.1. hypothesis 1: canadians will exhibit more raising than americans when measuring [aɪ] and [aʊ] raising, i expected to see more raising before voiceless consonants than voiced consonants having a lower f1 and a higher f2. among all speakers, the voiceless f1 mean is 636.3 hz (stdev = 6437.3) and the voiced f1 mean is 690.03 hz (stdev = 4404.5). this is as expected with a lower f1 before the voiceless consonant. t(df)=-4.1, p<0.0001 this p-value rejects the null hypothesis. among all speakers, the voiceless f2 mean is 1342.3 hz (stdev = 66810.7) and the voiced f2 mean is 1318.5 hz (stdev = 33491.1). this is as expected with a higher f2 before the voiceless consonant. t(df)=0.62, p=0.5 this p-value is greater than 0.05 and cannot reject the null hypothesis. in order to analyze the canadian speakers data and the american speakers data separately before comparing them, i measured the f1s and f2s individually before voiceless and voiced consonants. among canadian speakers, the voiceless f1 mean is 604.2 hz (stdev = 4382.9) and the voiced f1 mean is 689.9 hz (stdev = 5188.6). this is as expected with a lower f1 before the voiceless consonant. t(df)=-4.2, p=0.0004 this p-value rejects the null hypothesis. the voiceless f2 mean is 1392.9 hz (stdev = 93039.4) and the voiced f2 mean is 1363.2 hz (stdev = 56167.2). this is as expected with a higher f2 before the voiceless consonant. t(df)=0.4, p=0.7 this p-value is greater than 0.05 cannot reject the null hypothesis. among american speakers, the voiceless f1 mean is 660.9 hz (stdev = 6679.6) and the voiced f1 mean is 690.1 hz (stdev = 4058.5). this is expected with a lower f1 before the voiceless consonant. t(df)=-1.7, p=0.09 the voiceless f2 mean is 1303.3 hz (stdev = 44060.3) and the voiced f2 mean is 1284.9 hz (stdev = 15782.8). t(df)=0.5, p=0.6 both of these p-values are greater than 0.05 and cannot reject the null hypothesis. what this colorado research in linguistics, volume 25 (2021) 6 essentially means is that there is not a significant amount of [aɪ] and [aʊ] raising among american speakers. figure 1 finally, when comparing canadian and american speakers, i expected to see more raising exhibited by canadian speakers having a lower f1 and a higher f2 than american speakers. the voiceless f1 mean for canadian speakers is 604.2 hz (stdev = 4382.9) and for american speakers is 660.9 hz (stdev = 6679.6). this is as expected with a lower f1 exhibited by canadian speakers. t(df)=-4.4, p<0.0001 this p-value rejects the null hypothesis. the voiceless f2 mean for canadian speakers is 1392.9 hz (stdev = 93039.4) and for american speakers is 1303.3 hz (stdev = 44060.3). this is as expected with a higher f2 exhibited by canadian speakers. t(df)=1.9, p=0.06 this p-value is greater than 0.05 and cannot reject the null hypothesis. 3.2. hypothesis 2: athletes on canadian/northern cities (c/nc) teams will exhibit more raising than athletes on other teams in order to analyze the c/nc teams’ data and the other teams’ data separately before comparing them, i measured the f1s and f2s individually before voiceless and voiced consonants. among c/nc speakers, the voiceless f1 mean is 646.9 hz (stdev = 8320.3) and the voiced f1 mean is 686.1 hz (stdev = 5271.1). this is as expected with a lower f1 before the voiceless consonant. t(df)=-1.8, p=0.09 the voiceless f2 mean is 1370.7 hz (stdev = 74765) and the voiced a pilot study on voice-conditioned vowel raising comparison 7 f2 mean is 1324.5 hz (stdev = 22705.8). this is as expected with a higher f2 before the voiceless consonant. t(df)=0.9, p=0.4 both of these p-values are greater than 0.05 and cannot reject the null hypothesis. among other speakers, the voiceless f1 mean is 624 hz (stdev = 4093.6) and the voiced f1 mean is 692.6 hz (stdev = 4043.8). this is as expected with a lower f1 before the voiceless consonant. t(df)=-4.3, p<0.0001 the voiceless f2 mean is 1309.7 hz (stdev = 56754.7) and the voiced f2 mean is 1314.5 hz (stdev = 42133.9). this is not as expected because the f2 is lower before the voiceless consonant instead of higher. t(df)=-0.09, p=0.9 both of these p-values are greater than 0.05 and cannot reject the null hypothesis. figure 2 when comparing c/nc and other teams, i expected to see more raising exhibited by c/nc teams having a lower f1 and a higher f2 than other teams. the voiceless f1 mean for c/nc speakers is 646.9 hz (stdev = 8320.3) and for other speakers is 692.6 hz (stdev = 4043.8). this is as expected with a lower f1 for c/nc speakers. t(df)=-2.6, p=0.01 the voiceless f2 mean for c/nc speakers is 1370.7 hz (stdev = 74765) and for other speakers is 1314.5 hz (stdev = 42133.9). this is as expected with a higher f2 for c/nc speakers. t(df)=1.01, p=0.3 both of these p-values are greater than 0.05 and cannot reject the null hypothesis. colorado research in linguistics, volume 25 (2021) 8 hypothesis 2a: canadian athletes on canadian/northern cities (c/nc) teams will exhibit more raising than canadian athletes on other teams when comparing canadian speakers to each other, i expected to see more raising exhibited by the speakers on c/nc teams. the voiceless f1 mean for c/nc speakers is 583 hz (stdev = 5373.8) and for other speakers is 617.5 hz (stdev = 3425.7). this is as expected with a lower f1 for c/nc speakers. t(df)=-1.9, p=0.06 the voiceless f2 mean for c/nc speakers is 1489.4 hz (stdev = 127263.9) and for other speakers is 1332.3 hz (stdev = 64831.9). this is as expected with a higher f2 for c/nc speakers. t(df)=1.8, p=0.08 both of these p-values are greater than 0.05 and cannot reject the null hypothesis. figure 3 hypothesis 2b: american athletes on canadian/northern cities (c/nc) teams will exhibit more raising than american athletes on other teams when comparing american speakers to each other, i expected to see more raising exhibited by the speakers on c/nc teams. the voiceless f1 mean for c/nc speakers is 676.3 hz (stdev = 7021.5) and for other speakers is 632.8 hz (stdev = 5027.5). this is not as expected with a higher f1 for c/nc speakers. t(df)=2.4, p=0.02 this p-value rejects the null hypothesis. the voiceless f2 mean for c/nc speakers is 1316.3 hz (stdev = 43281.7) and for other speakers is 1279.3 hz (stdev = 46360.9). this is as expected with a higher f2 for c/nc speakers. t(df)=0.7, p=0.5 this p-value is greater than 0.05 and cannot reject the null hypothesis. a pilot study on voice-conditioned vowel raising comparison 9 figure 4 3.3. hypothesis 3: canadians will show a smaller difference based on team location than americans in order to answer this hypothesis, i decided to make a new table representing location divided by nationality, showing the average f1 and f2 for each group. then, i calculated the difference in f1 between the two locations, then the f2 as well. this left me with four calculations of difference which i ran through a t-test assuming unequal variances, as i did with all my other t-tests. the results of the t-tests and the numbers used in the calculations are included in table 2 below. in the end, i cannot be certain if this was precisely the right way to perform this calculation, but the t-test requires at least two data points for the range for each variable, so this is what i believed to be the solution. i am open to feedback on this and would like to rerun the tests if there is a better method. table 2. differences in f1 and f2 canadians americans ca/nc other difference ca/nc other difference f1 583 617.5 34.5 676.3 632.8 43.5 f2 1489.4 1332.3 157.1 1316.3 1279.3 37 colorado research in linguistics, volume 25 (2021) 10 the mean difference when calculating the f1 and f2 together for canadians is 95.8 hz (stdev = 7515.4) and for americans 40.25 hz (21.1). this is different from expected as i predicted the canadians would have a smaller difference and while they do have a smaller difference in f1, they have a larger difference in f2. t(df)=0.9, p=0.5 this p-value is greater than 0.05 and cannot reject the null hypothesis. figure 5 4. discussion, limitations, and future work i chose to only study two examples of canadian raising, where the [aʊ] and [aɪ] raise in the nucleus to [ʌʊ] and [ʌɪ]. however, there are other linguistic patterns attributed to a central canadian accent which would be interesting to study in future research, for example, influences on canadian english in the british columbia region by americans from the pacific northwest geographic region (washington and oregon) and vice versa. in choosing this subject, i knew i was choosing to make audio recordings from youtube videos and other videos rather than recording a person’s speech. additionally, i was recording connected speech, rather than words on a list or within carrier sentences. if either of these had been different it could have led to more clear formants and better measurements, and it may have influenced the results. indeed, after completing everything for this assignment, i went back to the data and deleted a couple of data points which had not been very good, and it did change the results slightly, especially some of those p-values which are 0.06 or otherwise very close to the 0.05 mark. in future work on this topic, i would conduct all of the tests with recordings captured live with a pilot study on voice-conditioned vowel raising comparison 11 appropriate technology and the samples done within carrier sentences. finally, simply having more data would make the results better in terms of generalizing it or knowing whether i had a true sample. two ways to address this shortcoming might be: (1) better consistency of words being measured and having more words since, for example, i had the fewest measurements for [aʊ] before voiced consonants; and (2) measuring more speakers to get results that could be better generalized. my research question around the cross-linguistic influence of place of origin and current residence in the hockey community of practice cannot be answered conclusively in this paper. more often than not, my results looked as i expected them to but then, upon running the t-tests, it would be revealed that the p-value could not reject the null hypothesis. additionally, the data on american speakers was less conclusive overall, by which i mean to say there were more results that were not as i expected at the outset, according to the hypotheses. i believe that the influence of northern cities vowel shift should have been better accounted for in this data with different considerations in the hypotheses. as stated in section 1, dailey-o’cain (1997) observed [aɪ] raising (part of ncvs) and also [aʊ] raising in the same region. so, more clarity in the hypotheses could have accounted for variation among the american athletes in my data whose place of origin in the northern cities geographic region influenced their speech uniquely. i have a follow up research question related to multilingual hockey players and foreign-born hockey players coming to canada and the united states as young adults to work on north american teams. first, i am interested in bilingual french canadians who speak french and english as their native languages who play exclusively in canada and the united states on teams which are primarily english speaking (there is only one team of thirty-one which uses french in public communications). second, i am curious about european athletes who have come to north america at some point in their professional development. there are a lot of factors to consider in both of these questions, for example, schooling in languages other than english, at what time did a player who is a non-native english speaker begin learning english, and for how long has the speaker lived in north america, primarily speaking english. a surface level analysis looking at the spectrograms for a swedish athlete who briefly played in ontario and currently plays in colorado (other team) shows raising in the [aʊ] diphthong before a voiceless consonant but not in the [aɪ] contexts. i wonder how this came about, how younger athletes who have not played on other teams as long as he has would compare, and how this compares with athletes in similar colorado research in linguistics, volume 25 (2021) 12 situations. for example, there is a strong tradition of finnish athletes in dallas, texas, and i wonder about the cross-linguistic influence of finnish with texan english on top of these questions about canadian english in the hockey community. the broader sociolinguistic question about identity related to occupation in the hockey community of practice needs more study. expectations around speech patterns, public perception of the athlete, camaraderie on their team, national pride, and age or stage of professional development are all involved consciously or subconsciously in an individual’s produced speech. the hypotheses considered in this paper and unanswered questions outlined in this section, along with their findings, would prove valuable in the field of sociophonetics. references bray, andrew. canadian features in the speech of american-born nhl players. american dialect society's annual meeting. jan. 2019. poster presentation. chambers, j. k. “canadian raising.” canadian journal of linguistics/revue canadienne de linguistique, vol. 18, no. 2, ed 1973, pp. 113–35. cambridge university press, doi:10.1017/s0008413100007350. dailey-o’cain, jennifer. “canadian raising in a midwestern u.s. city.” language variation and change, vol. 9, no. 1, mar. 1997, pp. 107–20. cambridge university press, doi:10.1017/s0954394500001812. a pilot study on voice-conditioned vowel raising comparison 13 endnotes 1 this project originated as a term paper for my phonetics course at the university of colorado boulder (december 2020). special thanks are owed to dr. rebecca scarborough who encouraged me to choose data based on my interests, and to lainey adams who helped compile and annotate data for this project. semiotics of peasants in transition: slovene villagers and their ethnic relatives in america (sound and meaning) book review irene portis-winner. semiotics of peasants in transition: slovene villagers and their ethnic relatives in america (sound and meaning). london: duke university press. 2002. 187 pages. isbn: 0-8223-2841-0. $22.95 us. reviewed by tamara grivičić in her book “semiotics of peasants in transition: slovene villagers and their ethnic relatives in america (sound and meaning),” irene portis-winner presents a significant semiotic study that, unlike many previous semiotic studies limited to the analysis of discourse alone, traverses multiple dimensions of cultural texts. portis-winner’s study comprises a three-decade-long fieldwork analysis of transnational and ethnic qualities binding two communities: that of the little peasant village of žerovnica, slovenia, and its emigrant population in cleveland, ohio and hibbing, minnesota. special attention is also awarded to the question of ethnicity and ethnographer’s or author’s voice in ethnographic studies. the study considers a broad range of polysemous, multi-vocal, and polyindexical values of cultural texts unbound by time-space continuum, which in turn prompt the author to redefine ethnicity as a dynamic entity not limited by “timeless essence” of individuals but rather free of eternal verity too often ascribed to societies. the complexity that defines ethnic culture and transnationalism is illustrated through a variety of cultural texts throughout the book. these texts range from: official to non-official history of the area and the villagers, everyday life, beliefs, traditions, economy, power and domination struggle, continuous revival and change of traditions and customs, and how they index the significance of signs. portis-winner’s study is empirical in nature because it employs a method that involves finding out what happens within a cultural text, rather then merely being told. the theme of lotman’s unconquerable boundary-crossing cultural hero is carried throughout the book as it is uncovered from personal interviews of reflexive narratives, and interpretive, double-voicing, accounts of the extended human sign. chapter 1 (3-27) provides a brief introduction to the economic, social, and geographic properties of žerovnica, as well as of its landscape, landmarks and inhabitants during the first fieldwork study in the 1960s. the question of inner versus outer (non-member) point of view immediately surfaces as the author warns that the immediate peaceful impression of a harmonious village and its inhabitants is positively deceptive. tension-ridden relations amongst villagers are discussed and traced to the communist rule and its goal to obliterate peasant autonomy and traditions that were considered a threat to the conglomerate whole. the chapter also informs of the pervasive hardship and exploitation of the colorado research in linguistics. june 2005. vol. 18, issue 1. boulder: university of colorado. © 2005 by tamara grivicic 1 grivic?ic?: semiotics of peasants in transition published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) peasants, as well as the imminent impact of global modernization on the village structure following the slovenia’s declaration of independence in 1991. the author's initial impression of a harmonious community changes after she has spent time within the ethnic community and has gained insight into their traditions and practices. portis-winner fervently argues that accuracy of an ethonographer’s research relies heavily upon his or her ability to become a quasimember of the group under investigation. she effectively accomplishes this task through a continuous exposure to a variety of ethnic texts, amongst others, modeling her conclusions after many member perspectives. i consider “semiotics of peasants in transition: slovene villagers and their ethnic relatives in america” a testament to the importance of efficient ethnographic work and applaud portis-winner’s efforts to provide us with such a valuable study. the last part of chapter 1 offers a taste of juxtaposition between the member-perceived vibrant and active life of the slovene emigrant community in cleveland, their clearly marked attachments to their slovene village, and the deteriorating, tension-ridden, and mistrustful community of žerovnica. an initial introduction is made to the changing semiotic aspects of objects and signs brought along by the migrants to the new world. portis-winner argues that semiotic changes, from practical to emotive and aesthetic, serve to reinforce the ethnic identity of slovene americans. part ii (28-74) comprises chapters 2 through 4. in this section, portiswinner provides a rich account of issues pertaining to traditional terminology of culture, society, nationalism, ethnic identity, and transnationalism, all relevant for the understanding of the study at hand. chapter 3 (43-49) is dedicated to a significant and recurring issue of non-member interpretation of cultural texts and modes of unearthing the communicative objects that are significant in the construction of an inner point of view. portis-winner warns about the problem of authorial interpretation of traditions and customs, their usage and changes. she advocates the inner point of view as essential in ethnographic research because it may have different realities and coherence, therefore rendering the unidimensional authorial view at best inaccurate and at worst overly simplistic. chapter 4 (50-74) offers a detailed overview and discussion of theoretical and practical issues pertaining to ethnographic studies over the decades. it spans views and attitudes of many semiotically-oriented scholars from saussure, peirce, the prague linguistic circle headed by jakobson, moscow-tartu school and bakhtin, to lotman and others. each subsection of the chapter introduces a new stance of one of the above-mentioned authors with respect to the analysis and attitudes toward cultural texts. special attention is afforded to the concepts of sign, symbol, and index; polysemy or mutlivocality of texts; everyday behavior; context (heteroglossia); perception and interpretation of history; as well as the undeniable role of power, which often forces cultural significance onto signs. portis-winner substantiates her synopsis with a much-needed critique of the semioticians’ attitudes, their respective problems or benefits toward a more wholesome ethnographic study. reader should be warned that previous 2 2 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/4 doi: https://doi.org/10.25810/4srm-pe36 book review: irene portis-winner knowledge and familiarity with the subject matter are indispensable in understanding of part ii. part iii (75-124) includes chapters 5 and 6. chapter 5 (77-105) revisits the economic, cultural and geographic landscape of the village. portis-winner offers a detailed account of historical events of the slovene people. she draws on official records, community’s view of future, and survey of cultural texts (recollections, beliefs, tales, myths, autobiographies, and changing beliefs and object meanings), to successfully extrapolate the inner point of view. chapter 5 concludes with an account of a changing value system in the peasant village. chapter 6 (106-124) discusses the immigrant community in cleveland, ohio, and should be of great interest to a linguist as it addresses the bilingual aspect of slovene american culture. the immigrant population shows great attachment to their mother tongue, which has undergone phonological, lexical, and grammatical changes under the influence of a new environment. portiswinner delineates member attitudes toward the slovene language over several migrant generations. much of slovene american communication is marked by code switching, especially within second generation immigrants. the third generation immigrants however are said to have initially shown embarrassment at their grandparents speaking slovenian, but later that there was some indication of the younger generation’s interest in the revival of the language. the rest of chapter 6 elaborates on the survival and upward movement of the slovene community. the success is ascribed to the traditional values the immigrants brought with them: stubbornness, ingenuity, hard work, loyalty to their tradition, generosity, discipline, honesty and responsibility toward family, kin, and country. portis-winner recounts several immigrant narratives, which, she persuasively argues, shed light on the ethnic culture as a part of a larger cultural context. the stories are significant in that they provide reference to the experience and points of view of the slovene migrant. the indication of transformation is present in a variety of signs, verbal and non-verbal, and may be evidenced in the meaning and significance change for the original signs, the change that points to the similarities and differences between one’s ethnic culture and the new environment. part iv (125-155) includes chapters 7 and 8. chapter 7 (127-151) surveys the major social and economic changes that bear heavily on the social and psychological state of the two communities in juxtaposition. portis-winner shows that global modernization has had an opposing impact on the elder generations between the two communities, while showing much affinity in the impact on the youth. changes in values and traditions between and within these ethnic communities serve to support portis-winner’s claim about the dynamic nature of ethnicity, boundaries of which are expanded, crossed, and re-evaluated on constant basis. by analogy, ethnic narrators in this study are seen as human signs indexing ethnicity -an intertextual and interwoven phenomenon that comprises a complexity of identities. the author further equates ethnic actors 3 3 grivic?ic?: semiotics of peasants in transition published by cu scholar, 2005 colorado research in linguistics, volume 18 (2005) 4 with actors in a theater, both of which, she claims, are able to move from one world to another and therefore become transfigured or transnational. portis-winner concludes the book in chapter 8 (153-155) with a discussion of polysemous and polyfunctional nature of cultural texts. the two main points that i have derived from her study can be encapsulated in the following thought: 1.) in order for ethnographic studies to be of value, ethnographers must work to unearth the inner point of view and formulate their conclusions after having considered a network of cultural texts, and 2.) every culture and its ethnic identity are amalgams of polysemous and polyfunctional properties of its cultural texts, which are dynamic in nature, and no view of ‘society’ holds permanently true across time and space. the way(s) in which a tradition is going to be impacted is unpredictable. some values and traditions may be maintained, others lost, and still many simply altered to reflect and adapt to the changes of the new environment and the new times. this book provides a comprehensive synopsis of the essential aspects of a thorough ethnographic study. portis-winner set out to conduct an empirical ethnographic fieldwork study, which in turn provided her with the necessary experience to help define criteria for a better way of conducting ethnographic research. she accomplishes this by intimately studying two related ethnic groups during a span of 30 years. the longitudinal study affords her a quasi-insider perspective of the ethnic group and provides access to invaluable ethnic sources. this is exactly the strength of her approach and only enhances our trust in her evaluation. because of its multi-disciplinary nature, “semiotics of peasants in transition: slovene villagers and their ethnic relatives in america” should be of interest to semioticians, ethnographers, as well as linguists and linguistic anthropologists. since the book is relatively non-technical, serving an interesting and comprehensive introduction to ethnographic study, as well as providing remarkable insight into the life of a particular ethnic group, i highly recommend it to any casual reader. tamara grivičić university of colorado department of linguistics 4 colorado research in linguistics, vol. 18 [2005] https://scholar.colorado.edu/cril/vol18/iss1/4 doi: https://doi.org/10.25810/4srm-pe36 colorado research in linguistics 6-2005 semiotics of peasants in transition: slovene villagers and their ethnic relatives in america (sound and meaning) tamara grivičić recommended citation against optional wh-movement microsoft word nilep-cril2021-proof_final.docx 1 reflection on my story with cril chad nilep nagoya university during a graduate student conference in the early aughts i was privileged to join a mentoring session with sociolinguist penelope eckert. after describing her shifts from studying older speakers in france (e.g. eckert 1980), to teens in michigan (eckert 1989), to children in california (eckert 1998), eckert mused that it is only in retrospect that we impose a coherent narrative on careers in research. thinking about my own career in this way, i see that cril has played an important part. my contributions to colorado research in linguistics during my graduate education have proved to be fundamental in my career in at least two ways. first, a paper published in cril – and not one that i expected to be influential at the time i wrote it – is the most cited paper on my cv. second, working as a member of the review board and later as journal editor was invaluable preparation for work that makes up a large share of my professional responsibilities today. as a graduate student, i was required to complete a “synthesis paper”, an overview and synthesis of some portion of the scholarly literature, before advancing to candidacy. i was advised to engage with theory, but not to use linguistic data in the paper. since i had never seen such a paper, i was at quite a loss regarding how to create one. i didn’t understand what it should contain, how it should look, or what sort of analysis i should undertake. the question, “what should i do?” from a student is at once rather too vague and potentially too weighty to allow faculty to provide specific advice and guidance. yet my own ignorance was too vast to allow me to construct more useful questions. to remedy my ignorance, i began asking slightly more senior students, those who had recently completed the requirement, to allow me to read their papers. this proved to be invaluable, as several people offered advice – or at a minimum, condolences – regarding the requirement. i found kristine stenzel’s work1 particularly helpful for, although her goals and her geographical concentration were quite different from mine (compare stenzel 2005), we did share several areas of interest. moreover, seeing a concrete example of the type allowed me to begin imagining how i could approach the task myself, both in terms of content and style. with this example in mind, i was able to write my own synthesis, a review of some of the literature on code switching. colorado research in linguistics, volume 25 (2021) 2 having benefited from the work of students ahead of me, i thought it only fair to share my own work with those coming after. i let colleagues know that i would be happy to share my own synthesis paper to any student who wanted it, and sent several people copies. around this same time, colorado research in linguistics was having a resurgence thanks to the work of alan boydell and adam hodges with david rood (cril 2004). following a six year hiatus, cril was being revitalized as an open access online journal, with a focus on working papers by graduate students, reviewed and edited by graduate students. i therefore submitted a revised version of my synthesis paper to the journal, and in 2006 it was published. that paper published in cril is my most successful work to date in terms of influence on other published work. according to google scholar, the paper has been cited more than 400 times, with many of the citing papers in turn cited dozens of times. it has also led to opportunities to contribute additional work on code switching (hall and nilep 2015; nilep 2020), to say nothing of the many requests to review other work on code switching and related topics. i am convinced that, whatever the merits of the subject matter and content of the paper, a contributing factor to the success of that publication is the fact that it was freely available as an online, open-access publication. although there is some controversy around the effect of open access on publication impact, some work suggests that open-access publications are at least similar to commercially published subscription journals in terms of “impact” as defined by those journals (björk & solomon 2012). other scholars argue that total citations are a better measure of actual contributions to scholarship, and that publications that are free to access tend to be cited more often than comparable pay-walled articles in the same field (harnad & brody 2004; piowar et al. 2018). in any case, i conclude based on anecdote that my work published in cril has had nothing but positive effect on my own career. my experiences with cril provided valuable experience not only as an author but also from the editorial side. in 2004 and 2005 as a member of cril’s editorial board i had my first experiences peer reviewing scholarly work, and discussing publication decisions with the editors and other board members. from 2006 to 2008 i served as the editor of colorado research in linguistics. these experiences provided valuable training for my current role in academic research and publishing. reviewing the work of other researchers is a major part of my work today. since 2010 i have worked with the nagoya university writing center, where i am currently designated associate reflection on my story with cril 3 professor in the institute of liberal arts and sciences. one of my key responsibilities as a member of the writing center faculty is to work with scholars across the university in order to help them publish their work in international journals. this involves consulting with scholars about both the logical and empirical content of their research, and the rhetorical and linguistic elements of their presentation of that research. as such, i regularly consult not only with scholars whose research interests are similar to mine, but also with those in very different fields from my own. again i find that my experiences with cril have prepared me well. the journal has published both work in linguistics as such, and interdisciplinary work related to fields across cognitive sciences, social sciences, anthropology, and letters. working with the journal and with departments across the university of colorado provided me a broad familiarity with various approaches to scholarship and scholarly writing and publishing. in my current position a broad (if sometimes shallow) understanding of diverse fields, as well as interdisciplinary and cross-disciplinary experience is a boon. in 2011 my nagoya university colleagues and i started a journal partially modeled on colorado research in linguistics, called nu ideas. nu ideas was edited by graduate students, as cril was during the time i was a member of its editorial board. also like cril, it published work from graduate students, as well as post-doctoral researchers, faculty, and other scholars affiliated with nagoya university. as a publication of the writing center, the journal accepted work in a broad range of disciplines, written in any of five languages (see e.g. nu ideas 2012). whereas cril is primarily a working papers journal, nu ideas published full working papers as well as shorter summaries, and fully vetted research papers in their final version. this required assistance from a broad range of (anonymous) reviewers, and yeoman effort from a hardy cohort of graduate student editors. between 2012 and 2018, nu ideas published nine issues including nearly 50 papers, thanks to the efforts of thirteen editors2 sharing various duties at different times. the 50th anniversary of colorado research in linguistics provides an opportunity to reflect on the past. when imposing a narrative on my own career to date, i have to give the journal, and its editors and contributors, a major role in that story. i hope that the future of my career in academic writing and publishing will continue to be as rewarding as my past with the university of colorado and with cril has been. colorado research in linguistics, volume 25 (2021) 4 references björk, bo-christer, and david solomon. 2012. open access versus subscription journals: a comparison of scientific impact. bmc medicine, 10 (1).1-10. doi: 10.1186/1741-7015-1073 cril. 2004. colorado research in linguistics 17. doi: 10.33011/cril.v17i online: https://journals.colorado.edu/index.php/cril/issue/view/cril17 eckert, penelope. 1980. the structure of a long–term phonological process: the back vowel chain shift in soulatan gascon. locating language in time and space, ed. by william labov, 179–219. new york: academic press. eckert, penelope. 1989. jocks and burnouts: social identity in the high school. new york: teachers college press. eckert, penelope. 1998. vowels and nail polish: the emergence of linguistic style in the preadolescent heterosexual marketplace. gender and belief systems: proceedings of the 1996 berkeley women and language conference. berkeley: berkeley women and language group. hall, kira, and chad nilep. code-switching, identity, and globalization. the handbook of discourse analysis, 2, ed. by deborah tannen, heidi e. hamilton, and deborah schiffrin, 597-619. wiley. doi: 10.1002/9781118584194 harnad, stevan, and tim brody. 2004. comparing the impact of open access (oa) vs. non-oa articles in the same journals. d-lib magazine 10 (6). online: http://www.dlib.org/dlib/ june04/harnad/06harnad.html nilep, chad. 2020. code-mixing and code-switching. international encyclopedia of linguistic anthropology, ed. by james stanlaw. wiley. doi: 10.1002/9781118786093.iela0056 nu ideas. 2012. nu ideas: nagoya university multidisciplinary journal 1. online: http://nuideas.ilas.nagoya-u.ac.jp/volume1/1-1-contents.html piwowar, heather, jason priem, vincent larivière, juan pablo alperin, lisa matthias, bree norlander, ashley farley, jevin west, and stefanie haustein. 2018. the state of oa: a large-scale analysis of the prevalence and impact of open access articles. peerj life & environment 6.e4375. online: https://peerj.com/articles/4375/ reflection on my story with cril 5 stenzel, kristine. 2005. multilingualism in the northwest amazon, revisited. memorias del congreso de idiomas indígenas de latinoamérica-ii. austin: university of texas. online: https://ailla.utexas.org/sites/default/files/documents/stenzel_cilla2_vaupes.pdf colorado research in linguistics, volume 25 (2021) 6 endnotes 1 stenzel’s paper was a critical review of the literature on community multilingualism in the vaupés region of amazonia. my own major interest at the time was related to the language behavior of sojourners from japan in the united states. both projects involved language contact phenomena. 2 they are, in alphabetical order: isabelle bilodeau, jasmina damjanovic, hsu peihsin, thomas kabara, sa kou, shylaja d. molli, kanako morita, chad musick, isabelle vea, wang qian ran, yabushita momoko, taeko yamada, and zhang lin. simon potter and i also served as guest editors for one issue in 2013. political news interviews: from a conversation analytic and critical discourse analytic perspective colorado research in linguistics. june 2012. vol. 23. boulder: university of colorado. © 2012 by nina jagtiani. political news interviews: from a conversation analytic and critical discourse analytic perspective nina jagtiani university of colorado at boulder this paper employs conversation analysis and critical discourse analysis to examine institutionalized interaction. the majority of the data is taken from american and british political news interviews. the focus is on the found similarities between the two approaches in regard to institutional discourse, viz. the formalistic features which both conversation analysts and critical discourse analysts investigate. it should be noticed, however, that there are still remaining differences between the two approaches, since critical discourse analysts, compared to conversation analysts, also focus on broader discursive issues of talk. 1. introduction the purpose of this paper is to examine two different approaches to discourse, which are conversational analysis (ca) and critical discourse analysis (cda), in regard to the nature and structure of televisual political news interviews. my decision to choose the frameworks of ca and cda for my analysis of political news interviews is motivated by the fact that both are sociological approaches, analyzing social structure and activity, and both have examined news interviews in a more or less thorough fashion. ca’s approach to research demonstrates that the analysis is grounded in the observable orientations that the interactants display when engaging in conversational interaction (clayman & maynard 1995). in contrast, cda is mainly concerned with the way social power and dominance are practiced and challenged through written and spoken discourse within a social and political context (van dijk 2001). 1 since cda is an eclectic approach, my analysis of political news interviews is by and large not based on prototypical studies, given that the focus is on interaction. there are, however, several conversation analysts who have combined studies of interaction and power relations in their work, e. g. hutchby (1992, 1996a, 1996b, 1999) and clark & pinch (1988). feminist researchers, e.g. kitzinger (2000), have drawn on ca as well to analyze the relationship between language, gender and sexuality. thus a ca analysis can show how inequalities are reproduced in discourse. it is important to point out that the treatment of power relations in ca is different from those in cda. in ca such relations are 1 i will not reproduce the debate between ca and cda here, since this would be beyond the scope of this paper. 1 jagtiani: political news interviews published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 2 constrained, observed and understood from participants’ communicative conduct in interaction. whereas in cda, the focus is on finding evidence of the operation of power relations in discourse (wooffitt 2005). for the purpose of this paper, political news interviews can be defined as question-and-answer exchanges between two or more participants, which often are confrontational and challenging in nature, since adversarial and competitive questions occur frequently (mullany 2002). the interaction is formal and institutionalized, produced for an overhearing audience that does not actively participate (heritage 1985; clayman & heritage 2002). discursive conflict is here defined as “… the interaction of interdependent people [can include two or more participants] who perceive opposition of goals, aims, and values, and who see the other party as potentially interfering with the realization of these goals” (putnam & poole 1987:552). 2 generally, from a conversation analytic perspective, verbal conflict is interactionally managed and designed (hutchby 1996a; clayman 2002). in cda, by contrast, discursive conflict can be defined based on semantic criteria. as van dijk (2001) claims, conflict is “… discursively sustained and reproduced by derogating, demonizing, and excluding the [o]thers from the community of [u]s, the [c]ivilized” (p. 362). political news interviews represent an institutional genre; such interaction is different from ordinary talk. ordinary talk is a form of interaction that is not restricted to a particular setting. it is comprised of conventions and practices applicable to numerous social goals; however institutional interaction involves restricted interactional rules (heritage 2004). it needs to be mentioned here that there can be cultural as well as social variations between political news interviews, such as differences between public service and commercial television channels (clayman & heritage 2002; lauerbach 2004). although it is commonly understood that critical discourse analysts, as well as conversation analysts, are interested in the analysis of interactional features regarding political news interviews, critical discourse analysts also focus on the discursive content of interaction to understand the construction of power (e.g., wodak 2007). for the sake of comparison, however, in my synthesis of the two approaches i discuss the similarities of their treatment of political news interviews, looking at their analyses of the formal and interactional features of this genre, since to date the extent to which both approaches overlap in critical areas has received little attention. in part 2 of this paper, i introduce the two different approaches of ca and cda along with the general importance of their major contributions. in part 3 i discuss political news interviews from a ca and cda perspective. i discuss strengths and weaknesses of each framework, but also their points of agreement in their respective analytical approaches to political news interviews. the concluding part presents my own synthesis of the above approaches. although it 2 there are cultural differences in how verbal conflict is conducted. for example, according to schiffrin (1984) conflict talk can be used to demonstrate sociability. 2 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/2 doi: https://doi.org/10.25810/yeqf-qc86 political news interviews: from a conversation analytic and critical discourse analytic perspective 3 appears at first glance that the two frameworks are very different, it has become obvious that despite such methodological differences, ca and cda show striking similarities in the analysis of political news interviews. 2. two approaches to discourse: conversation analysis and critical discourse analysis 2.1 what is conversation analysis? harvey sacks is considered to be the founder of ca, a qualitative sociological approach to the study of the organization of social interaction (levinson 1983; heritage 1984). 3 the basis for a conversation analytic approach is the ethnomethodological claim that the focus for analysis is on participants’ understanding of the ongoing interaction. analysts assumptions should therefore be discarded. in ca, participants’ own interpretations of talk shape their following contributions to the discourse (wooffitt 2005). consequently, context is the product of participants’ actions and therefore locally produced in the given interaction. 4 conversation analysts consider communication as a joint activity; they are interested in analyzing how such jointly organized interaction is produced (sacks 1984). thereby the focus is on natural occurring talk, such as formal vs. informal and institutional vs. personal interaction. sequences and speaking turns within sequences are the primary units of analysis (sidnell 2010). in other words, ca is mainly concerned with explaining how coherence and sequential organization in discourse is constructed and understood in order to identify systematic properties in talk (levinson 1983). 2.2 what is critical discourse analysis? critical discourse analysts study how “… social power abuse, dominance, and inequality are enacted, reproduced, and resisted by text and talk in the social and political context” (van dijk 2001:352). in their analyses they uncover how discourse discriminates against minority and powerless groups drawing from wider structural contexts (wooffitt 2005). the notion of power is here “… conceptualized both in terms of asymmetries between participants in discourse events, and in terms of unequal capacity to control how texts are produced, distributed and consumed … in particular sociocultural contexts” (fairclough 1995a:1). according to fairclough 3 heritage & roth (1995) argue that without quantitative evidence it is impossible to get reliable results from individual case studies about how participants orient to questions during interview conduct. 4 according to stubbe et al. (2003) there are more flexible versions of ca that include contextual and socio-cultural cues. 3 jagtiani: political news interviews published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 4 (1989), there are two aspects of the relationship between language and power. first, relations of power can structure and represent the social order of institutions or societies, thus powerful groups are able to establish and determine language use. second, there is the actual exercise of power, which means to constrain other people by using language. cda is considered to be a shared perspective rather than a school or a methodology (bell 1995; van dijk 2001). it adopts a political perspective that aims to have an effect on social practice and social relations by revealing takenfor-granted power relationships (titscher et al. 2000). by making such political interventions, analysts have preconceptions about their data (wooffitt 2005). the concept of ideology is, not surprisingly, an important factor within cda. ideologies can be defined as sets of beliefs that activate practices and attitudes maintaining inequality. since discourse includes ideological assumptions, such assumptions operate to preserve the interests of powerful groups (wodak et al. 2007). in cda the basic units of analysis are texts and the corresponding context is crucial. here, texts are not only written documents; they can also refer to oral discourse or visuals or even to a combination of these forms. different researchers working in the field of cda draw on eclectic and cross-disciplinary analytical techniques for their analysis. in the next section i will discuss the characteristics of political news interviews from a conversation analytic and critical discourse analytic perspective. 3. conversation analysis and critical discourse analysis applied to an institutionalized genre 3.1 political news interviews from a conversation analytic perspective in this section, i will discuss the following characteristics of political news interviews from a conversation analytic perspective: neutrality in political news interviews, the possibility of interviewees challenging the interview format and the conflictual potential of such interviews. heritage (1985) suggested early on that the analysis of political news interviews could be a central contribution to the general field of institutional talk. such institutionalized interaction, which influences the talk of both interviewer and interviewee, can be regarded in many ways as different from ordinary conversation (atkinson 1982; clayman 1991). everyday conversation is comprised of conventions and practices appropriate for various social goals, while institutional interaction involves restricted interactional rules (heritage 2004). generally, political news interviews can be regarded as question-andanswer sequences. in other words, they involve a normative turn-taking system that restricts participants to either asking questions or answering them (heritage 1985; greatbatch 1986, 1988; clayman 1988, 2010; schegloff 1988/89; heritage 4 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/2 doi: https://doi.org/10.25810/yeqf-qc86 political news interviews: from a conversation analytic and critical discourse analytic perspective 5 & greatbatch 1991; heritage & roth 1995). the interviewer’s conduct is influenced by the cooperativeness of the interviewees. usually, they collaborate with the interviewer by withholding a response until a question is completed, thereby confirming the neutrality of the turn (clayman 1988; heritage & greatbatch 1991; heritage & roth 1995). interviewers have the right to keep the floor until a question is produced. they can perform a range of actions, such as challenging or affiliating, but they have to be carried out in question format (clayman 2010). important to mention is that the interpretation of such political news interviews as either cooperative or confrontational is culturally variable and prone to social change (lauerbach 2004). 3.1.1 neutrality in political news interviews interviewers need to maintain a formally neutral stance while interacting with their guests (clayman 2002). 5 if they do abandon their role as questioners they use certain strategies to maintain a neutral stance (clayman 1988; heritage & greatbatch 1991). a frequent technique is to produce assessments on behalf of others. such statements can be used to focus on but not to align with an interviewee’s expressed position (heritage 1985; heritage & greatbatch 1991; clayman 2010), (clayman 1988:483): (1) excerpt 1 (nightline 7/22/85:17) 6 ((discussing violence among blacks in south africa)) ir:  reverend boesak let me pick up a point the ambassador made. what assurances can you give us that talks between moderates in that country will take place when it seems that any black leader who is willing to talk to the government is branded  as the ambassador said a collaborator and is then punished ie:  the ambassador has it wrong. it’s not the people … here, the interviewee rejects the assessment by rebutting the words of the same third party that the interviewer introduced previously. another procedure called “mitigating” (clayman 1988:487) is used when the interviewer produces an evaluative statement and mitigates its strength. s/he is able to express her/his own point of view and in doing so can minimize the importance of her/his opinion. such techniques allow the interviewer to be interactionally 5 since the late 1990s it could be argued that in some cases interviewers have abandoned their neutral stance in certain institutionalized frameworks, especially in the case of the american 24 hour news channels (cnn, fox, msnbc). 6 here, “ir” stands for “interviewer,” and “ie” for interviewee. 5 jagtiani: political news interviews published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 6 confrontational while remaining officially neutral (clayman 1988). a further technique is called “formulating” (heritage & watson 1980; heritage 1985; clayman 2010). formulations can be used to clarify, refocus or underline prior talk, as well as to cooperate or challenge interviewees’ statements. again, by using these formulations the interviewer can maintain a neutral stance. interviewees can also make use of question reformulations to avoid some aspect of an interviewer’s question. before providing an answer they can paraphrase the question that was asked. after reformulating interviewees continue talking, and such subsequent talk builds on the reformulation rather than on the original question (clayman 1993). as has become obvious, interviewers can do more than ask questions in the course of an interview. they often implement argumentative statements in their questions that can challenge interviewees’ positions. such evaluative statements are embedded within interrogatives, so that each complete turn can be regarded as a question, indicating a neutral stance. 7 however, interviewees can challenge interviewers’ neutral stance, showing that they are constrained by interviewees’ responses (clayman 1988; heritage & roth 1995). 3.1.2 interviewees can challenge the interview format interviewers have the rights to manage the introduction and organization of topics. generally, interviewees are not able to shift from one topic to another. however, there are instances where interviewees can challenge the normative question-and-answer format of the interview in order to control the discourse. one way to accomplish this is to talk about something else prior to answering an interviewer’s question. greatbatch (1986:443) calls this practice “pre-answer agenda shifting”. another practice, called “post-answer agenda shifts” (greatbatch 1986:443), allows interviewees to change the topic after answering an interviewer’s question. both shifts are always produced in combination with a response. also, they do not challenge the turn distribution rights of interviewers since interviewees do not speak out of turn (greatbatch 1986). interviewees can also control the topic of their talk by ignoring the focus that has been established by a previous question, meaning they do not produce an answer but talk about something else (greatbatch 1986). generally, instances where interviewees take a turn that is not a response to interviewers’ questions represent a violation of the normative question-and-answer sequence of interviews (haworth 2006). 3.1.3 conflictual potential of political news interviews 7 it needs to be mentioned here that not all interrogatives indicate a neutral and objective stance in the same way. some questions, like negative interrogatives, are often more partial than others (heritage 2002). 6 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/2 doi: https://doi.org/10.25810/yeqf-qc86 political news interviews: from a conversation analytic and critical discourse analytic perspective 7 clearly, political news interviews include a fair potential for conflict. in situations where interviewees talk before interviewers have introduced the actual question, the concept of the interview breaks down. in such circumstances, as schegloff (1988/89) noted of a 1988 interview with then vice-president george bush that quickly turned antagonistic, the “interview … turned into a confrontation” (p. 224). that is, when participants discard the regulations of political news interview interaction and start engaging in hostile talk, the interview structure is abandoned (schegloff 1988/89; heritage & greatbatch 1991). according to schegloff (1988/89) the transformation of an interview to confrontation consists of two parts; the institutionalized turn-taking system breaks down and competitive overlap occurs (schegloff 1988/89:231-232): 8 (2) excerpt 2 (bush/rather, c. 04:15) [ ] rather: but mr. vice president, you went ta israel in ] bush: [yes ] rather: . hhhh anda member of your own sta:ff mister craig fuller.((swallow/(0.5)) has verified. and so did the o:nly other man the:re. mister ni:r. mister amiron nir, .hh who’s the israeli’s .hh to:p anti terrorist man, bush: [ye: [s. rather: [.hh [those two men >were in a meeting with you an’ mister nir not once, < but three: times. three times. underscored with you that this was a straightout arms [fer hostages swap.] = .h h h ] = bush: [w h a t t h ey:: ] (.) were doing.] = rather: =now [how do youhow] do you reconc ] i have (sir)] bush: [read the memo ] read the memo.] what they::] were doing. rather: how: can you reconci:le that you were there < ‘more than’ and ‘less than’ symbols indicate that the talk is rushed < > used in the reverse order they indicate that the talk is slow < ‘less than’ symbol indicates that the following talk is started with a rush (( )) description of events 23 jagtiani: political news interviews published by cu scholar, 2012 colorado research in linguistics 6-2012 political news interviews: from a conversation analytic and critical discourse analytic perspective nina jagtiani recommended citation summary: a constructional approach to as far as np is concerned 1 gregory: a constructional approach to as far as np is concerned published by cu scholar, 1997 2 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/2 doi: https://doi.org/10.25810/ms8x-2f18 3 gregory: a constructional approach to as far as np is concerned published by cu scholar, 1997 4 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/2 doi: https://doi.org/10.25810/ms8x-2f18 5 gregory: a constructional approach to as far as np is concerned published by cu scholar, 1997 6 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/2 doi: https://doi.org/10.25810/ms8x-2f18 7 gregory: a constructional approach to as far as np is concerned published by cu scholar, 1997 8 colorado research in linguistics, vol. 15 [1997] https://scholar.colorado.edu/cril/vol15/iss1/2 doi: https://doi.org/10.25810/ms8x-2f18 9 gregory: a constructional approach to as far as np is concerned published by cu scholar, 1997 colorado research in linguistics 1997 a constructional approach to as far as np is concerned michelle gregory recommended citation tmp.1537637138.pdf.x7hma two patterns for conversational closings in instant message discourse colorado research in linguistics. june 2008. vol. 21. boulder: university of colorado. © 2008 by joshua raclaw. two patterns for conversational closings in instant message discourse joshua raclaw university of colorado this paper investigates the methods used by speakers to end conversations in instant message discourse. the analysis describes two distinct patterns of closing sequences – expanded archetype closings and partially automated closings – used to make a closing relevant to the interaction. the structure of these patterns are demonstrated to be reliant upon speaker orientation to various social and technological aspects of the medium, such as online presence and programcreated automated messages. the analysis concludes that the ways in which speakers close conversations are similar in structure to spoken closings in faceto-face interactions, though contoured specifically to the online medium in their application. 1. introduction research on the interactional aspects of talk occurring in computer-mediated environments has often focused on those communicative features that the medium has shaped in some way. within only the past few years, markman (2006) has shown that the persistent nature of conversation in certain formats of computermediated discourse (hereafter cmd) can affect speaker orientation to the turntaking mechanism, and schonfeldt and golato (2003) have demonstrated how the order that transmissions are received within synchronous formats of the medium may shape how speakers repair utterances or correct errors. earlier research by garcia and jacobs (1999) has discussed how the separation of the message composition and message transmission processes found in “quasi-synchronous” forms of cmd may affect both the allocation of speaker turns and various forms of error correction, while rintel et al. (2001) looked more narrowly at speaker orientation to automated messages in chatroom discourse during opening sequences. though this focus on the medium is arguably widespread throughout many studies of the sociolinguistic aspects of cmd, it is especially prevalent within work that makes use of research methodologies that scholars have traditionally used to study spoken interaction, such as the conversation analytic framework (hereafter ca; see e.g. sacks 1992) used in the work cited above.1 the majority of contemporary scholarship in this area has incorporated discussions of medium-dependent features alongside empirical analyses of data from online interactions. this has resulted in a growing field of research that 1 this is perhaps due to the somewhat marked status of computer-mediation within these types of studies. 1 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 2 might be more accurately described as informed by the theories and methods of ca rather than strict examples of them, as more “traditional” applications of the framework generally focus solely on speaker interaction rather than the nuances of the communicative mode (see rintel et al. 2001).2 while this type of analysis is increasingly prevalent within the contemporary cmd literature, there are still many interactional features of computer-mediated talk that research has yet to describe using this or any other framework, a number of which have already been widely documented within conversation analyses of spoken discourse. the present study is a discussion of one such aspect of online interactions, the conversational closing, which entails the various methods that speakers use to leave an interaction. the analysis looks specifically at closings within instant message (hereafter im) discourse rather than the chatroom styles of talk most prevalent within similar studies of cmd. through a discussion of speaker orientation to aspects of the medium, such as online presence and platformprovided automated messages, i identify two patterns for closing sequences. the first of these, the expanded archetype closing, closely follows the structure of spoken closings but contains features unique to the medium and exhibits a slightly different preference structure based on speaker accountability within the online sphere. the second and less common pattern, the partially automated closing, replaces what would be entire turns at talk in spoken closing sequences with features specific to the medium, such as automated messages. 2.spoken archetype closings as schegloff and sacks (1973) first illustrated, conversational closings are intimately tied to the larger system of turn-taking that speakers employ during talk-in-interaction, though the structures of closing sequences tend to follow distinct patterns that remain relatively constant. the most prominent of these patterns, described by button (1987) as the archetype closing, involves the exchange of two sets of adjacency pairs between speakers. an adjacency pair is any unit of conversation where two consecutive turns are exchanged from one speaker to another, and the content of each turn is pragmatically related to the other so that the first part of the pair generally invites the second, as would occur in a question-answer pair. in the case of the conversational closing, we are concerned with the pre-closing pair and the terminal exchange pair. within a telephone conversation, these might look like the exchange in excerpt 1. 2 one partial exception may be research on telephone conversations, especially work on opening sequences where speakers are shown to orient to the telephone ring as a summons. even in this body of work, however, the focus is largely placed on speaker interaction. 2 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 3 (1) excerpt 1 1 joshua: okay, 2 mom: okay. 3 joshua: bye mom. 4 mom: goodbye. as defined within the archetype closing, the pre-closing consists of a pair of topically-neutral utterances that effectively “passes” the speaker’s turn to the interlocutor. the first pair part of this action invites a temporary suspension of the turn-taking mechanism of talk, an otherwise infinite loop of sorts in which speakers either keep or exchange turns with one another based on the interactional relevance of such an action (see sacks, schegloff, and jefferson 1974). the second speaker has two choices at this point in the interaction. they may align with the suspension, as by responding with a similarly topically-neutral utterance, therefore shifting the frame of the interaction towards a closing. alternatively, they may disalign with the suspension by continuing the conversation, either through the introduction of a new topic or through anaphoric reference. it should be noted that either choice may occur regardless of whether the second speaker recognizes it as such. it is not intention that is strictly important within a conversation analytic discussion of the closing sequence, but rather whether speakers notably align or orient to various aspects of the talk by providing, or not providing, a preferred or otherwise “expected” response. in the case of alignment with the first pair part of the pre-closing, the suspension of the transition relevance occurring after the first turn makes relevant an end to the conversation as the speakers cease the exchange of turns that form the backbone of active conversation. this exchange is illustrated in line 1 of excerpt 1 where joshua’s “okay” serves as topically-neutral in relation to prior turns at talk. mom's orientation to line 1 as a pre-closing can be seen in her use of a similarly topically-neutral utterance as her response. this allows joshua to begin the remainder of the closing sequence on his next turn. following a successful pre-closing the first speaker may then initiate a terminal exchange pair, what one might commonly think of as an exchange of goodbyes. this part of the closing can be seen in lines 3 and 4 with joshua’s “bye mom” and mom’s subsequent “goodbye,” and in the case of the telephone call from which the data sample originated, is followed by both parties hanging up. the original research by schegloff and sacks notes that the format of the archetype closing may be expanded in any number of ways, and in practice this is often the case. for example, speakers commonly make use of additional pre-closing sequences prior to beginning a terminal exchange, and they may provide accounts of why they are leaving the conversation or arrangements for making future plans with their interlocutor. these elements will be covered further in the discussion of conversational closings within computer-mediated contexts. 3 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 4 3.closings and openings in cmd while conversational closings have received generous attention in work on spoken language, language and communication scholars have conducted little research on their occurrence in online interactions. oftentimes any mention of closings is done as an aside within a larger discussion of interactional strategies rather than as the focus of the work, such as werry's (1996:53) description of addressivity in “expressions of greeting and farewell” on internet relay chat (hereafter irc) or herring's (1996) and hård af segerstad's (2002) citations of variation in email salutations within discussions of formality in the medium. work by researchers such as vallis (1999) fills this gap to some degree with analyses that look specifically at the management of closings in cmd, though these too are typically part of a larger study of other aspects of the talk. for example, vallis draws on the ca methodology to describe the frequent use of accounts in closings within irc. in her discussion she attributes any lack of a “formal” closing sequence offered by speakers leaving the program to server problems or a general lack of involvement with the chat room. however, even these more focused findings are a minor aspect of a larger study of opening sequences, accountability and recognition work.3 a more substantial literature may be found on opening sequences in cmd, and these are often relevant to some degree to discussions of closing sequences based on interactional similarities between the two. rintel et al. (2001) make use of ca to provide descriptions of speaker orientation to automated messages provided by the irc program, neatly contrasting the ways that speakers manage the opening of conversations in chatroom environments with those in spoken discourse. within this discussion, rintel et al. introduce a structured description of opening sequences that progress from what they term channel entry phases. due to the structure of irc, these necessarily begin with automated messages sent to both the entering user and the other users of the chatroom to announce (or confirm) the new user's arrival into the chat. as rintel et al. demonstrate, speakers may notably orient to these messages by directly responding to them and, in doing so, shaping their opening sequences accordingly. as irc and numerous other synchronous formats of cmd also make use of automated messages when speakers leave the program, it is logical that speakers will similarly orient to them during closing sequences in ways similar to those described in the literature on openings. while currently unexplored, this orientation is perhaps hinted at in vallis's previous recognition that these automated messages may serve as the only closing announcement for speakers who do not make use of more “formal” closing sequences. however, the influence of these particular aspects of the medium will be different in practice when speakers leave, rather than enter, the discourse environment, as conversation will 3 though still a notable contribution to work on closings in cmd, vallis only devotes a paragraph of her article to this phenomenon. 4 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 5 already be in progress in the former case. additionally, as the one-on-one chat format of instant messages ensures that speakers do not consistently experience the channel entry phases (or channel exit phases) of multiple users as in irc, both entry and exit strategies will likely differ between the two formats of cmd. 4.data collection and methods the present analysis makes use of data taken from a corpus of conversations held using the aol instant messenger (hereafter aim) program. aim is a freely downloadable messaging program used widely in the usa (where the research occurred) and numerous other wired societies, and its use is especially prominent among teenage and college-aged youths. data collection occurred in 2005 and involved compiling the chatlogs of interactions held between a cohort of 17 undergraduates at the university of colorado. each participant contributed multiple conversations to the corpus, and there were a total of 58 conversations containing a closing sequence. (the term is expanded here from its original definition in the ca literature to include closings that make use of both userinitiated closings and, in certain circumstances, automated messages from the aim program). permission was obtained by each user prior to the analysis, and each individual's screen name, or handle, was anonymized. the gender of the speakers was retained, as were their ages, though no other demographics were noted. ten females and seven males contributed to the corpus, and all participants were aged 18-21. the use of chatlogs, the text file records of an interaction created through a program after the conversation has taken place, has both merits and faults when used as the basis for any type of linguistic analysis. for those making use of a framework like conversation analysis, which is notorious for its narrow attention to detail within the transcription of the talk under scrutiny, this issue is a prominent one. despite the historical use of detailed transcripts, the majority of research that turns the principles of ca to computer-mediated talk has instead used chatlogs as a source of data. notable exceptions to this include garcia and jacobs (1999) and markman (2006), who both developed transcription methods of their own based on video recordings of chat interactions. the use of original transcription practices was done out of sheer necessity, as the current set of transcription standards (such as jefferson's method) were clearly developed to capture spoken discourse. they often focus largely on capturing phonetic or extralinguistic aspects of the talk that are either missing completely, or represented differently, in text-based forms of discourse. the current transcription systems are also designed to capture gaps or silences within the discourse, as these typically occur in microseconds in spoken conversation. they are thus illsuited for recording not only pauses that may often occur in lengths of minutes or even hours, but especially for those gaps in the separate message composition and message transmission processes that are significant to the organization of the 5 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 6 interaction. this lack of a standard for text-based transcription is a notable impediment not only for allowing more analyses of transcribed data into research on cmd, but in performing work on the medium that can only be adequately accomplished with narrow transcription of the talk with details obtained by watching the conversation unfold from each user's perspective. however, as researchers such as rintel et al. (2001) have suggested, the use of narrow transcription may not be absolutely necessary for interactional analyses of cmd, even those making use of the ca framework. they defend the use of what they term single-point logs, so named because they capture the entirety of the conversation from a single computer, by arguing that more “naturally occurring” data can be obtained from speakers chatting in their homes rather than in the research setting necessary for video observation.4 issues of practicality also surface within this discussion, as somehow shuffling each user into a room together and subsequently transcribing the experience of each user is a difficult and certainly arduous prospect. while these caveats make immediate sense for work on formats such as irc where dozens of speakers may be conversing within a single channel at once, it remains applicable even when examining the one-onone interactions of formats like aim. this is due to the numerous aspects of im interaction that are unique to more “spontaneous” uses of the program rather than the necessarily planned conversations held within a lab. for example, while these interactions do not feature the frequent joining phases of multiples users that speakers on irc experience, they can be shaped significantly by the online presence of a user and their peers through the aim program's buddy list feature (this will be covered shortly in a description of the program). additionally, the unique type of online presence afforded by the use of an away message, as shown by baron et al. (2005), may also affect the course of a particular conversation. each of these features would be difficult, if not impossible, to witness or to capture if all speakers were in a lab setting.5 the use of chatlogs within the present analysis is motivated to some degree by all of these concerns, although the researcher remains open to, and has used in more recent collections of data, both methods for the study of im interaction. 5.transcription methods although the data samples used in the present analysis are largely taken directly from the chatlogs, any gaps or silences between message transmissions 4 current screen capture programs make the use of video cameras unnecessary for capturing online interactions, and arguably produce a better quality capture for producing the transcription. however, having each participant make use of these programs, which are often expensive, outside of a laboratory setting carries with it its own problems. 5 the unfortunate paradox here, of course, is that these features of online presence can only be adequately captured through a video capture of the interaction, and thus a middle ground that researchers have not yet hit upon is necessary. 6 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 7 have been manually inserted by the researcher. these take the form of numbers within parentheses representing the number of seconds between the transmission on that line and the one following. thus, the following excerpt would illustrate a 6-second gap between the transmissions on lines 1 and 2. (2) excerpt 2 1 elliott: you guys have your own blog. (6.0) 2 girlbot: it's precious each measurement in seconds also includes a decimal place followed by the number of deciseconds in each gap. this is meant to emulate the transcription systems of spoken discourse, and is used despite the fact that the timestamp from the chatlog only records gaps in 1-second intervals. however, as the present analysis contains data samples with gaps measuring over a minute in length, the use of the decisecond place holder is meant to prevent any confusion for readers familiar with ca transcription. measurements of gaps longer than a minute will thus record the minutes, seconds, and deciseconds of the gap (transcribed as m:s.ds), so that the following excerpt would illustrate a pause of 1 minute and 23 seconds between the transmissions on lines 1 and 2. (3) excerpt 3 1 metonym: so yeah that was kind of lame. (1:23.0) 2 metonym is away the transmission on line 2 of this excerpt also illustrates the automated message created by the aim program when a user sets an away message. this appears as the user's screen name without a colon, followed by the message “is away.” a similar message occurs when a user signs out of the program, and similarly appears as the user's screen name without a colon, followed by the message “has signed off.” all other lines of transmissions containing a screen name followed by a colon are utterances sent by that speaker to the instant message box. 6.features of the aim program the aim program allows for text-based discourse that is most accurately described as prototypically synchronous (raclaw 2006). that is, conversations occur within an environment designed for the near instantaneous transmission and reception of messages between speakers, though in actual use communication may occur asynchronously due to relatively large gaps between speaker 7 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 8 transmissions.6 this can be contrasted with a format designed for asynchronous communication, such as email, where speakers must first manually open threads before reading and then replying to transmissions rather than doing so instantaneously as they would while instant messaging. in practice, however, users may exchange emails rapidly enough so that it seems synchronous. the prototype descriptor thus allows for the potential disparity between design and use that exists in most text-based information and communication technologies. to use the aim service a speaker must first log in to the client, after which their screen name becomes visible to all users who have the entering user on their buddy list. the entering user's buddy list similarly shows the screen names of those acquaintances who are currently also signed in to the service, and uses graphic symbols attached to a user's name on the list to show whether a user has entered an idle state or has set an away message.7 an away message is any message that a user sets for automatic transmission to anyone attempting to contact the away user. a user that is away can receive and read messages from others, but if they respond to these, or message anyone else through the service, their away status is cancelled by the program. though baron et al. (2005) and nastri et al. (2006) have previously discussed their numerous social uses, away messages may be typically described as a courtesy when leaving the computer or making oneself otherwise unavailable to converse with others through the program. with this decade's recent explosion in the availability of broadband and other ubiquitous internet connections, users are increasingly leaving their accounts logged in to the aim service at all times rather than signing out of the program when they have finished using it. as baron (2004) has shown, this constant online presence is especially common among college students, the demographic of the present study. this presence plays a significant role in how these types of users leave, and likely how they open, conversations held through the program. 7. conversational closings in cmd the closing sequences of conversations from the im corpus follow two general patterns. the first of these is shaped similarly to the archetype closing described earlier, though it is often expanded through multiple sequences rather than single exchanges of pre-closings and terminal exchanges. these expansions typically contain features such as accounts, arrangements, prefaces, hedges, or palliatives; their use within closings will be discussed shortly. with few notable 6 however, this is far from a clear distinction. the aim program is shaped in many ways to allow for synchronous interactions, through features such as the ability to set an away message or shift into an idle state are more likely oriented towards asynchronous communication. 7 a user enters an idle state when they have been inactive in the program for a set amount of time, typically 10 minutes. each user has the option of making their idle status public or private in regard to whether it shows on the buddy lists of others. 8 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 9 differences, the expanded archetype sequences closely follow the models for conversation closings described in analyses of spoken discourse. the second pattern involves the use of automated messages provided by the aim program as part of the closing sequence. these often occur within the position of the terminal exchange, appearing after a single pre-closing sequence or after a significant pause or silence. while these partially automated closings follow the basic structure of the archetype closing to some degree, their interactional uses are notably different from expanded archetype sequences occurring in both spoken and computer-mediated forms of talk. 8.structures of expanded archetype closing sequences expanded archetype closings begin with an utterance designed to serve two primary functions: it must alert one's interlocutor of the speaker's intention to close, and it must shift the frame of the conversation towards the termination of the interaction. both occur by introducing a first pair part of a pre-closing sequence that makes the idea of closing somehow relevant to the conversation. in the majority of im conversations, this was accomplished through the use of accounts, explanations or justifications for why the speaker initiated the closing. these were often followed by arrangements, or plans made with the other speaker to talk again at a later date. these sequences are then followed by the terminal exchange between both speakers, after which a user typically sets an away message or signs out of the program.8 this action will be referred to here as the trigger, as it generates an automated message both within the instant message box and on the buddy list, and its position within the discourse as the post-closing. the following excerpt demonstrates an extended archetype closing sequence that makes use of these sequences and features. here, metonym brings the conversation to a close after a long stretch of talk with pudding. (4) excerpt 4 1 metonym: so i should like, probably start writing my paper (11.0) 2 pudding: yeah i should probably go to bed (8.0) 3 metonym: so i will talk to you tomorrow, jah [yes]? (7.0) 4 pudding: jah [yes] (6.0) 5 pudding: good luck writing!!! (2.0) 6 metonym: thanks! (2.0) 7 pudding: latahz [i'll talk to you later] (3.0) 8 pudding: haha, bye (9.0) 9 metonym is away 8 while it is common for both speakers to eventually set an away message or sign out of the program at some point after an interaction, only the first of these actions is considered here to be interactionally relevant as subsequent triggers occur after the first user has already made themselves in some way “unavailable.” 9 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 10 metonym begins the closing by providing an account of why he intends to leave. while his utterance is not topically neutral, as pre-closings are described to be in the archetype closing model, the account still makes the closing relevant to pudding by introducing another activity that cannot be done while continuing the conversation. the confirmation of this relevancy can be seen in pudding's orientation to and alignment with metonym's pre-closing in line 2, where she provides an account of why she has to leave as well. the exchange of accounts is followed by an arrangement by metonym and an agreement by pudding in lines 3 and 4. this can be described as a general arrangement, since it does not specify a specific time or location for the future interaction to occur. among the larger collection of im conversations, general arrangements were as common as more specific arrangements, though specific arrangements universally involved meetings occurring offline rather than through aim or some other computermediated environment. an example of the latter type of arrangement can be seen in excerpt 5 below. (5) excerpt 5 1 girlbot: hey babes ive got to eat (13.0) 2 fingers: i'm pretty hungry too (5.0) 3 fingers: maybe go out for pizza (7.0) 4 girlbot: okay :) (2.0) 5 girlbot: meet me here for brek [breakfast]?(5.0) 6 fingers: yeah def [definitely] (6.0) 7 girlbot: okay see yaaaa (3.0) 8 fingers: see ya!! (27.0) 9 fingers has signed off this preference for offline arrangements may be attributed to a growing preference for cmd to serve as a supplement to other forms of interaction rather than serve as a singular vehicle for social relationships, as in squires's (2003) analysis of what she has termed multimedia relationships. in excerpt 4, following the speakers' exchange of the general arrangement and the anaphoric references to the initial account in lines 5 and 6, the terminal exchange occurs in lines 7 and 8. metonym sets his away message a short time after this sequence, an action that “finalizes” the closing beyond the terminal exchange by framing him as unavailable for future conversation. setting an away message in these contexts may thus be considered a final turn at talk, as it shares information with one's interlocutor that is potentially interactionally relevant, similarly to how a transmission sent directly from the user might. although this relevance will be more closely examined within partially automated closing sequences, the ability of an away message to convey a sense of unavailability to a speaker is still notable within the expanded archetype closing. as long as pudding's message box was still open when metonym set the away message, she 10 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 11 received the automated message seen in line 9 and could interpret it as metonym no longer being available as an interactant within the conversation. in lieu of setting an away message a user may sign out of the program, an action that is interactionally relevant in ways similar to the setting of an away message (though in the latter case the user is slightly “more” available as they may still receive messages from other users). alternately, a user can elect to leave the conversation without performing either action. of these possibilities, the setting of an away message as a post-closing occurred most often, happening in 44 of the 58 conversations in the corpus. signing out of the program occurred in 12 of the remaining conversations, and using neither trigger occurred in only 2 of the conversations. the preference for either of the former options likely lies in the accountability that comes with the continuous online presence afforded by the aim program. when a speaker sets an away message or signs off, it is generally to let other users know that they are not immediately available to talk.9 thus, a speaker signed in to the program without an away message may be viewed as conversationally ready, an appearance that speakers who have just closed a conversation often may not seek to convey. the general preference in the postclosing for setting away messages rather than signing out of the program is likely due to the previously noted increase in ubiquitous internet connections, but may also stem from a desire to be somewhat more available to other speakers (as users can still receive messages while their status is set to away) than signing out would allow. post-closings were used far more frequently by the speaker initiating a closing sequence, occurring in 39 of the 56 closings within the corpus. in cases where second speakers initiated post-closings, there was typically a larger gap separating the terminal exchange from the post-closing than when first-speakers made use of them, as can be seen in line 8 of excerpt 5 where there is a pause of 27 seconds between finger's second pair part of the terminal exchange and the point at which he signs out of the program. it is notably more difficult to measure “significant” silences objectively in im discourse than in spoken discourse, as pauses of a minute or longer may be the norm during certain exchanges within the former mode while pauses of a second or longer in the latter are typically worth noting. this difficulty is of specific concern when using a framework, such as ca, that attributes so much potential meaning to a gap between speaker turns. the present discussion thus considers the possible significance of larger silences within an interaction based on their contrast to gaps between prior turns at talk. the 27second pause occurring in excerpt 5 is thus considered notable due to the range of pauses from 2 to 13 seconds during the remainder of the closing sequence. the tendency for larger gaps to occur between the actions of lines 8 and 9, and in similar actions occurring throughout the corpus, is perhaps due to an expectation that the speaker closing the conversation is the one who will also leave the 9 this is, of course, a generalization to some degree; as baron et al. (2005) discuss, away messages can also serve as invitations to speak rather than messages of strict unavailability. 11 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 12 program. if this is the case, second speakers may then wait for the first speaker to use a post-closing before they themselves leave the program. the greater tendency overall for first speakers to use a post-closing, however, is more likely due to the fact that second users may simply not be finished with the program at the end of the conversation. in addition to the accounts and arrangements common to expanded archetype closings, speakers also made frequent use of features such as prefaces, hedges, and palliatives. prefaces and hedges are markers that index uncertainty to some degree, such as um, well, and maybe. palliatives are any portion of an utterance used to show appreciation or apology, often to soften a blow of some sort, such as a refusal in you're a really nice guy, but no thanks. an example of these features within online closings can be found in excerpt 6. (6) excerpt 6 1 fishfood: so like, i love you and all, but i should probably start 2 my homework :/ (9.0) 3 granola: blech, thats stupid (13.0) 4 fishfood: haha homework is stupid (5.) 4 granola: yet makes you unstupid (3.0) 5 granola: or does it (5.0) 6 fishfood: haha (3.0) 7 fishfood: okay, i'll see you tomrrow (6.0) 8 granola: ok see you then (3.0) 9 fishfood: later! (2.0) 10 granola: byeeeeeeeee 11 fishfood is away fishfood begins the pre-closing sequence in lines 1 and 2 of this excerpt with the preface “so like” followed by a common palliative structure, a clause containing a positive utterance, here a compliment, that is contrasted with what follows in the next clause using the conjunction but. what is contrasted here is the account, and thus the termination of the conversation that is indexed by the account's function as a pre-closing. the account itself additionally contains hedging with the inclusion of “i should probably” before the announcement of the specific account activity, and the use of a “slanted mouth” emoticon (:/) that may be interpreted as conveying dissatisfaction or sadness. granola's response in line 3 is in alignment with fishfood's initial framing of the pre-closing activity as something negative, as it potentially insults both the act of doing homework and the initiation of the closing sequence. what the response does not do, however, is clearly align with fishfood's initial attempt at closing in lines 1 and 2, and it is not until fishfood's use of an arrangement in line 7 that granola is seen to orient to the closing during his response in line 8. returning briefly to the matter of accounts and arrangements, excerpt 6 demonstrates that second speakers do not universally orient to accounts as pre-closings. it might then be argued that one reason for the 12 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 13 frequent use of both features together within a closing is to provide multiple chances to secure the second speaker's alignment with the act of closing. another potential function of using both features in closing sequences lies in the conversation analytic notion of preference (e.g. pomerantz 1984; sacks 1987). within adjacency pairs, such as the question-answer pair cited earlier, recipients of the first pair part of the sequence are generally limited to sets of responses that will be pragmatically relevant or otherwise make sense within the context of the interaction. posing a question to one's interlocutor thus generally limits their response to a sensible answer, if they know one, or a clear acknowledgment if they do not. these second pair parts may also be significantly shaped by whether they align or disalign with the first pair part. those in alignment are, with few exceptions, considered to be preferred, while those in disalignment are dispreferred. this conceptualization of preference does not refer to the more common understanding of psychological preference within a response, but rather the relationship between the various parts of a conversational sequence. thus, whether a first speaker truly desires an answer to their question is irrelevant. as providing an answer aligns with the earlier act of asking the question, it is thus a preferred response to the first pair part. as an alignment with a first pair part is generally framed to be somewhat “expected,” they are typically unmarked in their use. dispreferred actions or utterances, however, are marked by elaboration or explanation, long gaps or pauses between the first pair and second pair parts, and features that serve to mitigate their dispreferred nature. structurally, these mitigations often include the addition of features such as prefaces, hedges, palliatives, and accounts (schegloff 2007). returning again to the frequent use of accounts and arrangements in im preclosings, there is the possibility that their inclusion is intended to mark some dispreferred action within the discourse. researchers such as cameron (2001) and coppock (2005) have suggested that the frequent inclusion of these features within closing sequences indicates that speakers may view closing sequences in general as somewhat dispreferred. this attachment of dispreference is discussed by both sources as a reaction to the possible face threat to one's interlocutor that comes with ending a conversation. because of this, it may be the obligation of the first speaker to assure the second that they are free of any fault for the first desiring to leave. the account accomplishes this action by stating the reasons behind the closing, while the arrangement allows the first speaker to ensure their interlocutor that future contact is desirable, thus saving face for the second speaker. due to the frequent use of not only accounts and arrangements, but numerous other marks of dispreference such as prefaces, hedges, and palliatives, it is likely that there is another motivating factor for closings in im discourse to be oriented to as if dispreferred. for example, it is possible that the continuous online presence afforded users of the aim program also gives the impression of a continuous availability for interactions. to end a conversation is to directly remove oneself from this availability, an action that may thus be viewed as 13 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 14 dispreferred. this idea of accountability in opting out of continuous availability may be supported by the practice of returning accounts with other accounts, as discussed earlier. this is seen in excerpts 4, where metonym initiates the preclosing by stating that he has to begin writing a paper, and pudding aligns with his account by admitting that she needs to go to sleep. both speakers imply that they need to leave the conversation, which thus absolves the first speaker of any responsibility for leaving the conversation by framing this as a mutual necessity. considering the initiation of closing sequences to carry accountability also explains the first exchange in excerpt 6, where granola provides a negative assessment of fishfood's pre-closing. the type of explicit reaction seen here, where speakers outwardly critique or object to a closing, only occurred in two interactions, however. it was far more common among the interactions in the corpus for speakers to make use of the more subtle markers of dispreference seen in previous examples. 9.the structure of partially automated closing sequences the second pattern of closing sequences observed in the im interactions may be described as partially automated closings, so named because the automated message provided by a trigger serves a central role in the closing sequence. this may be contrasted with the use of triggered messages as post-closings in expanded archetype closings, as their use in these sequences can be better described as “supplementing” the closing sequence already provided by the speakers. within partially automated closings, these automated messages can be more accurately seen as replacing, either in part or as a whole, the exchange of turns leading to the termination of the interaction. as with other closings, these sequences first require an action that invites a suspension of the turn-taking mechanism and allows the closing to be seen as relevant to the conversation. these were consistently accomplished through the occurrence of one of two sequences throughout the corpus. in the majority of partially automated closings, first speakers made use of a pre-closing sequence similar to those seen in earlier examples, often employing an account to shift the conversation towards its termination. in the interactions shown in excerpt 7 and excerpt 8, neither speaker initiates a concrete pre-closing sequence, but rather a significant pause occurs between the last exchange of turns and the use of the trigger. (7) excerpt 7 1 sonorant: hey i have to go shower before i go out tonight. (5.0) 2 prettygirl: okay. (3.0) 3 sonorant is away 14 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 15 (8) excerpt 8 1 leetdood: hey, i should probably go to bed. (11.0) 2 paperdoll: sweet dreams, hun (8.0) 3 leetdood has signed off as in examples of expanded archetype closings, accounts were frequently used within first pair parts to introduce an action or responsibility that necessitated the end of the interaction. however, in each example within the corpus these accounts stood on their own as the pre-closing rather than being expanded through multiple sequences. in excerpts 7 and 8, line 1 presents an account from the first speaker that the second speaker aligns with in line 2. in line 3 the first speaker then initiates a trigger that serves as the final turn of the interaction, after which the conversation ends. the second speaker's alignment with the first part of the pre-closing consistently appeared throughout the partially automated closings, and each occurred prior to the trigger. however, given the somewhat unsure nature of the organization of utterances in most synchronous forms of cmd (e.g. garcia and jacobs 1999; markman 2006), it is certainly possible for the second pair part of the pre-closing to fall after the trigger within these sequences. additionally, it is possible that the alignment with the first pair part may not occur at all, though this is likely to be infrequent given the inherent preference for receiving a response to one's utterance. the choice of whether to set an away message or to sign out of the program in the final turn of these interactions is likely affected by the previous notions of individual internet connection and online presence, though further motivation may occur in the nature of the account provided within the pre-closing. in excerpt 7, sonorant's account still leaves him potentially available to talk later that night, either after his shower but prior to going out or after returning from going out. his use of an away message in this context is thus in alignment with this potential for future interactions that night. conversely, leetdood's account in excerpt 8 implies that he will be asleep for the remainder of the night and therefore unavailable to talk. in signing out of the program, leetdood's choice of trigger similarly aligns with the unavailability that his account provides. similar comparisons could be drawn between the remaining examples of pre-closings from the corpus. for example, one speaker set an away message when leaving the computer to eat, while another signed out of the program when leaving for work. the argument for consistent alignment with the form of trigger and the nature of the account is notably reliant upon the multimodal relationships (squires 2003) that the collegeaged speakers likely held, and it should thus be noted that they may not be applicable to all interactions. as excerpts 7 and 8 illustrate, using a trigger is a viable final turn in this pattern of closing. however, within these examples it is difficult to place the role of the trigger within the closing sequence. though it occurs after a pre-closing and serves to end an action, similar to the terminal exchange seen in expanded archetype closings, the trigger is organized asymmetrically rather than as part of 15 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 16 an adjacency pair. as excerpt 9 illustrates, however, speakers may also orient to the triggered message as if it was a first pair part of a terminal exchange. (9) excerpt 9 1 wicket: damn im gonne [gonna] be late to class! (6.0) 2 element: haha aight [all right] (3.0) 3 wicket is away (8.0) 4 element: see ya 5 auto-response by wicket: in class :( [wicket's away message] in line 1 of the excerpt, wicket first begins the pre-closing with an account that element aligns with in line 2. wicket sets her away message shortly after this alignment during line 3, and element provides a goodbye during line 4. his utterance initiates the transmission of wicket's away message in line 5, which is designated as such by the aim program using the preface “auto-response by wicket.” due to the organization of the interaction, it is possible to interpret the trigger as falling under either the pre-closing or the terminal exchange. the former interpretation is plausible only if element's goodbye is seen as the first pair part of the terminal exchange, a reading that also requires wicket's automated away message in line 5 to serve as the second pair part. however, it is difficult to assign interactional relevancy to this type of message as speakers do not appear to orient to them within these types of closing sequences. further, even in doing so, this interpretation would require the message in line 5 to be interactionally relevant in the present excerpt but not in excerpt 7 or 8 where it does not appear. it would also require that element would orient to the auto-response in line 5 as a second pair part to the terminal closing despite the fact that it was not sent from wicket. finally, it is also statistically unlikely that element's utterance in line 4 served as the first pair part to the terminal exchange as he was the second speaker within the original pre-closing. the majority of first pair parts of terminal exchanges were initiated by the first speaker throughout the corpus, with exceptions occurring in all but three closings due to significant gaps of over 20 seconds between the exchange and the final part to the pre-closing. a silence of this nature did not occur between lines 3 and 4 of the present excerpt. it is therefore unlikely that element's use of a goodbye in line 4 is simply his initiation of a terminal exchange. it is more plausible for the triggers in excerpts 7, 8, and 9 to be interpreted as a terminal transmission that may or may not be taken up by the second speaker as a terminal exchange. as excerpts 7 and 8 both illustrate, the second speaker is not obligated to return the terminal transmission, and this is likely because the first speaker has made himself in some way unavailable to talk through the trigger and has thus removed himself from the turn-taking mechanism. however, the option to orient to the terminal transmission as the first pair part of a terminal exchange is still viable, and thus the use of a trigger in 16 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 17 partially automated closings may be seen as akin to the role of the terminal exchange in expanded archetype closings. the nature of the accounts used in pre-closings occurring prior to terminal transmissions may help to further explain their viability in these closing sequences. in many accounts the first speaker implies a sense of urgency in the explanation of why they have to leave the interaction. these often take the form of an immediacy account, such as in excerpt 9 where wicket needs to terminate the interaction because he will otherwise be late to class, and excerpt 10 below. here, fingers provides an urgent account explaining how he needs to visit the library before they close. (10) excerpt 10 1 fingers: shit ive got to get to the lirary [library] like now befor 2 they close 3 girlbot: that sucks! its freezing out!!! 4 fingers is away 5 girlbot: call me when you're back! both excerpt 9 and excerpt 10 express their urgency through the use of curses (damn, shit) as prefaces to the accounts, while excerpt 9 additionally features an exclamation point at the end of the account and excerpt 10 features the temporal reference “now” in all capital letters (perhaps implying it to be read as if yelled). the sense of urgency often seen in these pre-closings may explain why there is only one sequence exchanged prior to the terminal transmission rather than the multiple sequences seen in other forms of closings. it may also explain the lack of features such as hedges within the accounts, as providing some uncertainty as to whether the first speaker truly needs to leave would likely detract from the sense of immediacy being conveyed. partially automated closings do not follow the “full” structure of closing sequences that is typically conceptualized in the expanded archetype sequence. therefore, in creating a sense of urgency speakers can mitigate the use of this shortened sequence, as well as the use of an automated message, to conclude an interaction. the use of a terminal transmission was also evident in conversations that did not otherwise appear to contain a pre-closing or other recognizable portion of a closing sequence. within this type of interaction the terminal transmission occurred directly after a significant gap in the discourse, such as the over six minute pause occurring in excerpt 11 (below) between fishfood's utterance in line 5 and his trigger in line 6. 17 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 18 (11) excerpt 11 1 granola: i'm the indie kid who hangs out and listens to deathcab 2 (8.0) 3 fishfood: i liked postal service better (12.0) 4 granola: blech (5.0) 5 fishfood: haha (6:25.0) 6 fishfood is away like the other closing sequences discussed in this analysis, this pattern of the partially automated closing requires the suspension of the turn-taking mechanism in order to make a closing relevant to the interaction. however, this is accomplished here through a literal cessation of turn-taking rather than through the pre-sequences used in every other closing sequence. rather than describing the interaction as featuring no distinct closing, it is more accurate to describe the silence between lines 5 and 6 as serving the role of the pre-closing, as the use of a more formal pre-sequence is rendered unnecessary due to the already present suspension of the turn-taking mechanism. moreover, the trigger in line 6 is best described as the final turn that terminates the interaction. this interpretation remains valid regardless of whether the speaker intentionally used the trigger to close the interaction, as it sends a message to the second speaker that the conversation can no longer continue. this is due to the previously cited focus in ca on speaker orientation rather than intention. additionally, as previous examples of both expanded archetype and partially automated closings have shown, speakers typically orient to triggers as if they were part of the closing sequence, and this is what occurs here. as stated earlier, this pattern of partially automated closings occurred far less frequently than those containing a more concrete pre-sequence, and excerpt 11 was in fact its only occurrence within the corpus. however, the structure used in this example has occurred with relative frequency in interactions that the researcher has both participated in and observed casually. the closing is thus included here and discussed as a viable component of im discourse. 10.conclusion this article has examined two patterns of closing sequences available to users within im discourse. the expanded archetype sequence discussed here has been shown to closely follow the structure of closings found in spoken discourse, but also makes use of a post-closing sequence in the form of a trigger and often contains various markers of dispreference outside of the accounts and arrangements typically found in spoken closings. these markers were attributed to the possible accountability of a user to retain a specific type of availability in the constant online presence afforded by the program. future work on closings in other formats of cmd would be valuable in seeing whether this accountability 18 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 19 carries over into other types of interactions. the partially automated sequence introduced here shows that automated messages can serve a more focal role within a closing sequence than those seen in the expanded archetype sequences, and that concrete pre-closing sequences may not be necessary for an interaction when prior gaps in the interaction have already suspended the turn-taking mechanism. these sequences also demonstrate the interactional relevancy that medium-specific aspects may hold for speakers in cmd. like rintel et al. (2001) have shown with opening sequences, these sequences show how speakers may orient to these aspects in ways that uniquely affect the sequential organization of the talk. finally, because speakers were shown to orient in numerous ways to various aspects of the medium throughout the discussion, this work serves to encourage future work to continue the tradition of examining the effects of the medium on interaction in conjunction with analyses of how speakers react to these medium-specific features throughout the discourse. references baron, naomi. 2004. “see you online: gender issues in college student use of instant messaging.” journal of language and social psychology 23(4): 397423. baron, naomi, lauren squires, sara tench, and marshall thompson. 2005. “tethered or mobile? use of away messages in instant messaging by american college students.” in r. ling and p. pedersen (eds.) mobile communications: re-negotiation of the social sphere, 293-311. london: springer. button, graham. 1987. “moving out of closings”. in g. button and j. lee (eds.) talk and social organization, 101-151. multilingual matters ltd. cameron, deborah. 2001. working with spoken discourse. sage, london. coppock, elizabeth. 2005. politeness strategies in conversation closings. stanford university: unpublished manuscript. garcia, angela, and janet jacobs. 1999. “the eyes of the beholder: understanding the turn-taking system in quasi-synchronous computer-mediated communication.” research on language and social interaction 32(4): 337367. hård af segerstad, ylva. 2002. “use and adaptation of written language to the conditions of computer-mediated communication.” phd diss., department of linguistics, göteborg university. 19 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 20 herring, susan 1996. “two variants of an electronic message schema.” in s. herring (ed.) computer-mediated communication: linguistic, social, and cross-cultural perspectives, 81-106. amsterdam: john benjamins. markman, kris. 2006. “following the thread: turn organization in computermediated chat.” paper presented at the annual meeting of the national communication association, san antonio, tx. november 2006. nastri, jacqueline, jorge pena, and jeffrey t. hancock. 2006. “the construction of away messages: a speech act analysis.” journal of computer-mediated communication 11(4). http://jcmc.indiana.edu/vol11/issue4/nastri.html pomerantz, anne. 1984. “agreeing and disagreeing with assessments: some features of preferred/dispreferred turn shapes.” in j. atkinson and j. heritage (eds.) structures of social action: studies in conversation analysis, 57-101. cambridge: cambridge university press. raclaw, joshua. 2006. “punctuation as social action: the ellipses as a discourse marker in computer-mediated communication.” paper presented at the annual meeting of the berkeley linguistics society, university of california at berkeley. february 2006. rintel, e. sean., joan mulholland, and jeffrey pittam. 2001. “first things first: internet relay chat openings.” journal of computer-mediated communication 4(4). http://jcmc.indiana.edu/vol6/issue3/rintel.html. sacks, harvey. 1987. “on the preferences for agreement and contiguity in sequences in conversation.” in button & lee (eds.), talk and social organization, multilingual matters, clevedon, pp. 54-69. sacks, harvey. 1992. lectures on conversation. oxford: basil blackwell. sacks, harvey, emanuel a. schegloff and gail jefferson. 1974. “a simplest systematics for the organization of turn-taking for conversation.” language 50(4): 696-735 schegloff, emanuel a. 2007. sequence organization in interaction: a primer in conversation analysis. cambridge: cambridge university press. schegloff, emanuel and harvey sacks.1973. “opening up closings.” semiotica 8(4): 290-327. schönfeldt, j., and a. golato. 2003. repair in chats: a conversation analytic approach.” research on language and social interaction 36(3): 241-284. 20 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/4 doi: https://doi.org/10.25810/5e6r-fk53 two patterns for conversational closings 21 squires, lauren. 2003. “college students in multimedia relationships: choosing, using, and fusing communication technologies.” american university tesol working papers, 2: 1-18. vallis, rhyll. 1999. “members' methods for entering and leaving #ircbar: a conversation analytic study of internet relay chat.” in k. chalmers, s. bogitini and p. renshaw (eds.) educational research in new times, 117-27. flaxton: post pressed. werry, christopher c. 1996. “linguistic and interactional features of internet relay chat.” in s. herring (ed.) computermediated communication: linguistic, social, and cross-cultural perspectives, 47-63. amsterdam: john benjamins. 21 raclaw: two patterns for conversational closings in instant message discourse published by cu scholar, 2008 colorado research in linguistics 6-2008 two patterns for conversational closings in instant message discourse joshua raclaw recommended citation microsoft word cril_paper_raclaw.doc 1. introduction in recent years, korean pop music has become a major force in the music industry, gaining a considerable fandom in the process. based mainly on the internet, the k-pop fandom has become a complex network of fans on various social media sites ravenously consuming and discussing the genre in question. as the fandom has grown and changed, different archetypes of fans have been identified and commented upon. one such archetype of fan is the “koreaboo”. viewed wholly negatively, the archetype of the “koreaboo” is one that is seen as being obsessed with and appropriative of korean culture and the korean language by other fans. in particular, the linguistic habits of fans viewed as koreaboos have come under fire. this koreaboo register is characterized by using korean words in otherwise non-korean writing. in this paper i argue that other fans have created what i am calling “mock koreaboo” in order to express their negative view of koreaboo fans and their linguistic habits. following in the tradition of mock spanish (hill 1995; hill 1998), i show that mock koreaboo is a form of mock language being used by k-pop fans to mock other k-pop fans. like mock spanish, fans use this mock koreaboo to disparage koreaboo fans through an exaggerated version of koreaboo speech. while mock koreaboo uses korean linguistic resources to do this, it does not ultimately aim to mock korean people, unlike mock spanish. instead, mock koreaboo seeks to criticize the appropriation of korean by those viewed as koreaboos and the ideology associated with this archetype. in addition, i argue that mock koreaboo uses stancetaking to show superiority over fans seen as koreaboos. mock koreaboo utilizes elitist stances (jaworski & thurlow 2009) to negatively evaluate the figure of the koreaboo and elevate the user. through its mocking nature, fans can use mock koreaboo to show how they reject the ideologies of the koreaboo and are a better fan as a result. i will also discuss how koreaboo speech and mock koreaboo create several orders of indexicality (silverstein 2003). with korean as the nth order, koreaboo speech and mock koreaboo constitute n+1st and n+2nd orders of indexicality respectively, creatively building off each other. in this paper, i will discuss the archetype of the koreaboo fan and how k-pop fans react to koreaboos through the use of a mock koreaboo register. i will discuss the use of this register, the attitudes mock koreaboo represent, as well as how it utilizes stance and indexicality. 2. background: k-pop, fans, and global flows k-pop is, at its core, simply pop music that comes from south korea. the history of modern k-pop reaches back to the 1990s, and throughout this history the genre has always been heavily influenced by music by black americans (mosely & mcmahon 2020). while k-pop has always been global, in the past ten years k-pop has become increasingly globalized in its reach via fan efforts and increased marketing outside of asia (chang & park 2018; herman 2019). this has resulted in fan communities springing up all over the world, participating in mainly online contexts (swan 2018). the online experience of being k-pop fan is significant to understanding k-pop fandom in general. one study of k-pop “stan” twitter has categorized it as a community of practice (malik & haidar 2020). malik and haidar interview and analyze interactions between a multinational group of fans of the k-pop group monsta x. through their research, they show that these fans have formed strong bonds with each other over their shared interests despite their distance and linguistic differences. chang & park view the online experience of bts fans as a modern “tribe” despite physical and cultural distance between the fans themselves (2018). while such research shows the structure of k-pop fans online, there is little research into the ideology of k-pop fandom and its linguistic behaviors as a whole. research on individual fans such as those who post reaction videos on youtube exists, looking into topics such as polyculturalism in fans’ blending of korean pop culture with black culture (oh 2017). another study based on interviews with canadian k-pop fans found that these fans experience k-pop as a hybrid cultural product that allows them to participate in globalization (yoon 2018). the phenomenon of k-pop’s international reach and fanbase owes itself to the now global travelling of information. linguistic, visual and cultural information now “flow” around the world, finding their ways into different places and along with their meanings and uses. scholars like h. samy alim have researched these global flows of culture. alim (2009) specifically has looked extensively into how hip-hop culture has spread around the world, becoming a new site for identification and meaning making. one specific case of this in linguistics literature is the korean pop star cl, who, in one of her music videos, has adopted signifiers of hip-hop and chola culture to represent her being a “bad girl” (garza 2021). embodying the nappeun gijibae character through the visual language and linguistic resources of chola, blackness, and hip-hop, cl can come across as a “bad bitch” but also reinforces negative stereotypes about these groups. in linguistics, another noted case of korean media utilizing cultural flows is the case of the portrayal of korean-american. in the 1990s, the style of the korean-american, especially those from los angeles became in vogue. the visual language of k-pop stars from the us, heavily influenced by hip-hop itself, the english language, and the representation of la came to represent coolness (lo & chi kim 2012). the figure of the korean-american was seen as a skilled-multilingual, especially compared to koreans who viewed themselves as bad at english. eventually korean-americans lost their status, replaced by the “transnational korean returnee” figure. 2.1. stancetaking and indexicality in du bois seminal work on stance, he defines stance as a “public act by a social actor, achieved dialogically through overt communicative means, of simultaneously evaluating objects, positioning subjects (self and others), and aligning with other subjects, with respect to any salient dimension of the sociocultural field” (2007: 163). in ochs’ view, stance is an act that can index further meanings. through her writing on the indexicality of gender, ochs’ positions stance as an intermediatory step for language to index social meanings such as gender (1992: 342). people use language to create stances through direct indexicality, and then these stances indirectly index social meanings. many types of stances exist including epistemic and deontic stances, but this paper will focus on elitist stances. jaworski and thurlow’s work on this topic sees elitism as a person “making a claim to exclusivity, superiority, and/or distinctiveness” that needs constant upkeep (2009: 196). through their study of travelogues in british newspapers, jaworski and thurlow show that while elitist stances are still an evaluation of a subject, they do so by making this claim to superiority. elitist stances are enacting certain ideologies wherein the object that the subject is evaluating is better than others, and therefore the subject is also better than other subjects. in ochs’ view, this would be an example of indirect indexicality where one indexes their own superiority through making these stances of elitism about other subjects. the identity of elitism is itself evaluative and is upkept through the evaluative power of stancetaking. as i will show later, k-pop fans exhibit their superiority to the fans they consider to be koreaboos through the use of elitist stancetaking in mock koreaboo. indexicality is a wide-reaching concept in sociocultural linguistics, and a very important one. in a simplified way, indexicality is the way social meanings are constituted through language (ochs 1992). these meanings can be anything from identities to activities or positions on something. one of the most influential writings on indexicality is michael silverstein’s work on indexical order (2003). in this article, silverstein presents the concept indexical order, a system where indexical meanings can be created on top of each other incorporating the previous meaning but innovating on it at the same time. the first indexical meaning is called the nth order of indexicality and any meaning that follows is therefore n+1st order of indexicality. in this example, the nth order is the “standard” pronunciation of new york english. lower middle-class people were shown to hypercorrect their speech, in this case having higher rates of /th/ and final /r/ than even upper middle-class people. this hypercorrection of pronunciation constitutes a n+1st order, according to silverstein. in comparison to ochs’ view on indexicality, silverstein’s indexical order also shows how indexical meanings are layered. while ochs describes direct and indirect indexicality as a way that meaning can be constructed through different resources, indexical order shows how indexical meanings can be utilized and innovated upon to create even more new meanings. blommaert (2007) innovates on this concept by conceptualizing indexicality as polycentric. in this view, indexicality revolves around “centres” comprised of “complexes of norms and perceived appropriateness criteria, in effect the larger social and cultural body of authority” (p. 118). “centres” can be anything of any size, from a singular person to a global concept. conversations and interactions can and often are polycentric, moving between centres of indexicality. blommaert gives the example of a rasta dj in south africa changing his speaking patterns based on how he wants to be perceived and what the topic is. koreaboos and other kinds of k-pop fans can be seen as orienting towards a korean centre of indexicality, particularly through the way that koreaboos utilize korean linguistic resources as is shown later in this article. mock koreaboo, in contrast, is then orienting towards a non-korean centre through criticizing koreaboo’s linguistic behaviors. as i will show in this article, mock koreaboo utilizes elitist stances to create a separate indexical order from koreaboo speech, and to index their disapproval of the figure of the koreaboo. 3. methods this paper is based in an online ethnography of k-pop fans. digital etthnography, also called virtual ethonography or cyber-ethnography, is simply ethnography that is done in virtual spaces. the methods of digital ethnography can be as wide and varied as traditional, in-person ethnography. beneito-montagut (2011) argues that digital ethnography is not different than traditional ethography, it just requires different “research decisions regarding the particular field and location of research” (p. 719). in this paper, beneito-montagut describes a method that follows a user throughout their online life and activities, just as a researcher might follow a person to their job and hobbies. hallet and barber (2014) advocate for incorporating someone’s online life into the rest of those experiences, as those cannot be meaningfully separated anymore. bonilla and rosa (2015) put forth the idea of hashtag ethnography, wherein a hashtag on social media can itself become a site of research, linking various ideas and posts on a certain topic together in one place. like beneito-montagut (2011), bonilla and rosa also argue for the following of specific users to get the context of their posts. in this project , i have gathered data in two main ways: searching and browsing. for searching, i looked through mainly k-pop specific subreddits (r/kpoprants, r/kpopthoughts, etc.) and twitter. i searched for specific terms related to this project such as “oppar”, “unnir”, “koreaboo” and other terms, specifically korean words which are associated with koreaboo speech. in regards to browsing, i collected any relevant posts which i found while browsing these spaces on my own time, which i often do. collected posts were stored in a spreadsheet along with the time of posting, theme of post, and the platform it was found on. 4. data and analysis: what is a koreaboo? the origin of the term “koreaboo” is an amalgamation of the term “weeaboo” and “korea”. “weeaboo” itself is a term decribing people not of japanese descent who are so into japanese media and culture athat they denounce their own in favor of japanese culture, or those who wish they are japanese but not (ewens 2017; hidayat & hidayat 2020). the word itself is a nonsense word, coming from a the comic “the perry bible fellowship”, in which a person is attacked for merely mentioning the word “weeaboo” (birney & keogh). the term is pejorative, and people rarely self-identify with “weeaboo”. koreaboo is along the same lines. in one reddit post on the subreddit r/kpoprants, a poster defines a koreaboo as “someone who is obsessed with korean culture so much they denounce their own culture and call themselves korean” (engenie 2021). a commenter on this post replies that this is too strong of a definition, saying that it is someone who becomes obsessed with korea, but does not interact with the wider culture of korea other than k-pop or k-dramas. like “weeaboo”, koreaboo is also only applied to people who are not of korean descent. apart from this post, the koreaboo is mainly defined through actions, particularly through the use of the korean language in otherwise non-korean contexts, especially by nonkorean speakers. another reddit post on the same forum asks the question, “is my friend a koreaboo?” sorry for any korean words spelt wrong. i really need to know cause i don’t know how to tell her. her daily vocabulary is mixed with broken korean all the time, when i call her she answers her phone with a “yeoboseyo”. but here’s where it’s starts getting cringe for me she calls her mom “eomma” and speaks to her whole family in korean... like if her siblings asks her questions or just conversing in general she’d be like “ani” “ne” “jinjja” “aish” “wae” “omo” “aigoo” and “hajima”she says aigoo all the time. i asked her one time do they know what she’s saying and she told me they’re starting to catch on mind you we are african american. she throws in little phrases of korean here and there and calls her husband oppa and he’s not even korean he’s puerto rican. one time she asked him to help her with this video game she was stuck on and she started to do aegyo to get him to help her and she did it in this whiny voice cringe alert oppaaaa~~ jinjja pouts help me with this game and proceeded to laugh and tell me she does aegyo when she doesn’t get her way. she also talks to her regular friends in korean too they don’t know a thing about kpop or korea at all.... she talks them and they’re always like what did you say or what does that mean and it gives me second hand embarrassment soooo bad. last thing we both went to see bts in chicago at soldier field what coincidence our lyft driver was a older korean man and the whole way she kept talking in broken korean and i’m in the back like because we’re actually in close proximity of someone who actually speaks korean. i love her so much but i don’t want to hurt her feelings so i’m just gonna post it on here. what do y’all think? figure 1: reddit post by stanbts_ (stanbts_ 2021) the poster describes the behavior mainly as using korean words such as “yeoboseyo” [hello] to answer the phone instead of the english “hello” and talking to her family in korean despite them being african americans living in chicago, according to the poster. the comments on this post universally categorize this person as a koreaboo. one commenter recommends that the friend learn korean formally, suggesting that the bad attempts of using korean are the problem here. another commenter worries that they are a koreaboo since they tend to use korean but are studying the language. another user replies that “it’s normal since it comes from you actually intently learning the language but koreaboos just want to say annyeong yeorobeun [hello everyone] because it sounds cool.” this again suggests that improperly using korean without the proper intent is the issue with koreaboos. an article on what signs to look out for in a koreaboo also focuses on language (napper 2019). interestingly, this article includes using korean romanization but not the korean writing system, known as hangul. this also adds to the inauthenticity in language dimension, where using hangul is seen as appropriate, but using romanization is unacceptable. according to posts like this and my own experience, some ideologies associated with koreaboo speech are: description: example: gendered terms of address oppa, unnie, hyung, noona “i love you” in korean, especially the word “love” without the verb ending saranghae [i love you], saranghaeyo [i love you], i sarang [love] you using romanization instead of hangul saranghae [i love you] instead of 사랑해 korean intensifiers jinjja [very], neomu [very] exclamations omo, aigoo, aish other basic vocabulary ani/aniyo [no], ne [yes], wae [why] using korean words in an otherwise english utterance i neomu [very] like this table 1: features associated with koreaboo speech the chief characteristic associated with koreaboo speech is the use of korean words and other linguistic resources outside the context of a korean utterance. one major aspect of the above examples is the mention of the word “oppa”. the korean word “oppa”, written 오빠 in hangul, is a gendered term of address used by women to address a slightly older man that one is close to, like a friend or partner (jeong & yu 2021). in south korea, female fans often use “oppa” to address the k-pop boy groups they are a fan of, likely drawing on the romantic partner aspect of the term (tracy wonwoo 2022). non-korean fans have picked up on this, and some nonkorean fans have begun to use the term as well, in much the same way. in the figure 1, the friend is said to use the term with her husband, who is also not korean. the article makes reference to the use of this and the similar term “unnie” (written 언니 in korean) used by women to address slightly older women, calling it “pure k-boo culture.” searching twitter, i am able to find many instances of korean used in an otherwise nonkorean context, as is considered koreaboo speech. consider the following excerpt of a tweet: “jinjja me는 freaking out” (ella 2022). in this tweet, the user uses both a korean intensifier written using romanization jinjja [really] and the korean topic particle 는 written in hangul attached to the english me. the rest of the tweet is in english, with the only other korean terms being the names of people and a korean television program. the use here seems to mainly be for aesthetic purposes, or to index the user’s like of korean culture. according to this user’s personal website linked in their twitter profile, she is from the philippines and only speaks basic korean, showing that this is not a native speaker mixing two languages that she is fluent in. as evidenced by figure 1 and the comments of that post, this type of usage is considered to be koreaboo speech, wherein korean is used in a superfluous manner and out of context. 5. mock koreaboo since being a koreaboo is seen as undesirable by other k-pop fans, a form of mock language called mock koreaboo has been created. the concept of mock language was pioneered by jane hill in her article on “junk spanish” (1995). while she later switched the wording to “mock spanish”, in this article she defines “junk spanish” as “a set of strategies for incorporating spanish loan words into english in order to produce a jocular or pejorative key” (p. 205). key aspects of mock language in this article are the pejoration of the mocked language, adoption of lexical and morphological material in order to do this pejoration, and the exaggerated mispronunciation of the mocked language. mock spanish is also part of a white american “light” register, where the meaning of spanish material is completely stripped to create a nonserious variety of language for white americans. while mock spanish can uphold racist ideas, this is not always the case for mock languages. chun (2004) investigates korean-american comedian margaret cho’s use of “mock asian and finds that cho’s use of mock asian is “legitimate”, showing that cho uses the register not only to position herself in opposition to asianness, but also to be critical of mock asian as a practice. her use of mock asian can de-center whiteness and critique racist imaginings of asian women. another example of this intra-ethnic mockery appears again in lo and chi kim (2012). the article describes a korean comedy skit where korean-americans are in a korean class, speaking korean with bad grammar, heavy american accents, and plenty of english. while not described as so in the article, this sketch could be seen as an example of mock korean-american, making fun of their poor command of their heritage language to platform their own savvy in languages and cultural references. mock koreaboo can be found all around social media, including twitter and a parody meme subreddit called r/kpoopheads. an example of mock koreaboo is below (translations of korean and korean-derived mock koreaboo in brackets): chingoose [friends], i just learnt that joy unnie [gendered term of address] has started dating crush flop oppar [gendered term of address] me neomu [very] sed now, ottokhae [what do i do]?!1 me sarang [love] joy unnie since me 2 year old, me thought joy unnie [gendered term of address] sarang [love] me bacc even tho me onli 11 i thought i felt something, but now i realised there was never any sarang [love] anyone else go through such betrayal? ottokhae [what do i do]?? me sarang [love] joy unnie [gendered term of address] but me get cheated on me go doxx unnie [gendered term of address] and crush flop oppar [gendered term of address] now, onion chingoose [bye friends] figure 2: a reddit post containing mock koreaboo (outrageous-bottle-72 2021) in this passage, the poster writes as a koreaboo character who is upset that a k-pop artist the poster likes has announced she is in a relationship. prior to this post, k-pop singer joy has announced her relationship to the k-pop singer crush, respectfully referred to using the gendered terms of address unnie and oppar. here, the poster mocks the attitude that some k-pop fans have regarding their favorite singers dating by pretending to be jealous of the relationship and betrayed by the singer joy. jealousy is a common theme among mock koreaboo, mocking the perceived intense attachment to k-pop idols that koreaboos have.1 mock koreaboo is similar to koreaboo speech, but differs in several key ways. one of the most relevant ways it differs is the blatant and intentional misspellings. as seen figure 2, the misspellings are not only of korean, but english as well. this can be seen in the term onion/onion haseyo which is used for the korean greeting annyeong [hi] or annyeong haseyo [hello]. in the above example we can also see chingoose, a deliberate misspelling of chingu, meaning friend in korean. this term may also be mocking the out of context application of the english plural -s to the word, becoming chingus. some of the english misspellings include bacc and sed. description: example: deliberate misspellings of both korean and english chingoose [mock form chingu meaning friend], onion haseyo [mock form of annyeong haseyo meaning hello], delulu (delusional) mock gendered terms of address oppar, unnir use of “sarang” as if an english verb i sarang [love] you high emoji use (common in mocking speech on the internet in general) framing an unusual person as target of affection lee soo man oppar (lee soo man is the former ceo of sm entertainment) table 2: features of mock koreaboo one of the most salient examples of mock koreaboo is shown through oppar. an intentional misspelling of oppa [gendered term of address], oppar directly mocks the overuse of the term by koreaboo fans. the reason for the added -r on the end is not currently known, but may be because rhotic r is not found in korean, marking “infelicitous, anglicized pronunciation” of the term (jeong & yu 2021: 832). a related form, unnir also exists as a mock form of unnie [gendered term of address]. oppar is one of the key signs of mock koreaboo, is used by other kpop fans even without using mock koreaboo. this, along with the themes of jealousy, seems to mock the attachment of fans to k-pop performers, specifically how they are “delusional” in thinking that k-pop artists are their romantic partners. this is especially true when oppar is used with an unusual person, such as entertainment company ceos like lee soo man, jy park, and bang si hyuk as company ceos usually do not have a fanbase themselves.2 like mock spanish, mock koreaboo is both pejorative as well as used in a humorous context (hill 1995). the difference here is that instead of trying to mock a group of people for speaking their native language, mock koreaboo seeks to mock people for using a non-native language inappropriately. instead of being racist itself, mock koreaboo is possibly trying to make fun of perceived racism through misusing the korean language. koreaboos are often seen as appropriative of korean, using the language in inappropriate circumstances. through exaggerating this appropriation, users of mock koreaboo can jokingly criticize the appropriation itself. mock koreaboo has a more limited use case, as well. hill presents examples in all forms of media such as radio and birthday cards. mock koreaboo is limited more to directly mocking koreaboo fans and is mainly found on social media. as compared to chun’s analysis of margaret cho’s mock asian, mock koreaboo is not necessarily a reclamation. the ethnic background of those who use mock koreaboo is unknown, but it does not seem to be restricted to korean or asian people. it can be seen to fight racism in a similar way, but koreaboo speech is not used with racist intent as mock asian is usually used. mock koreaboo seems to be more of a way to distinguish oneself from undesirable fans. 5.1. elitist stancetaking and indexical orders in mock koreaboo mock koreaboo is taking an elitist stance against koreaboo fans. as jaworski and thurlow define it, elitism is making a claim to superiority based on whatever moral claim the speaker wants to distinguish themself (2009). elitist stances work through evaluation, where the subject evaluates an object either positively or negatively, and implicitly evaluates another subject negatively, elevating the subject. an obvious example of this is the following tweet: all the delulu sasaengs crying : but oppar you looked at me and fell in love didn’t you?! : as i said i didn’t make eye contact with armys3 for 2 years now….. (@sgtcurrypants 2021) in this tweet, the poster critiques “delusional stalker fans” who believe that a member of bts fell in love with them. the tweet calls these fans “delusional stalker fans” delulu sasaengs where delulu is short for “delusional” and sasaeng is a korean term for an obsessive fan. some terms, like sasaeng and maknae which refers to the youngest member in a group or family, are considered acceptable uses of korean by fans. i am not sure why this is at this time, but it is an option for further analysis. in the third line, the tiger emoji represents bts member v, who in the video attached to the tweet spoke about not being able to see his fans. here we see the poster using oppar as a clear marker of mock koreaboo, and the fans in this fake dialogue being portrayed as unreasonable through the attributed text in all caps and the content of the statement. the attributed text to v also portrays the fans as being unreasonable by directly dismissing them, treating them as ignorant. by portraying these “sasaeng” or stalker fans in this way, the poster is evaluating these fans as less than and positioning themselves as a better type of fan. they are distinguishing themselves from the fans that are obsessive and too into it. while this tweet is explicit in its elitist stance, mock koreaboo itself is imbued with elitist stances. in figure 2, mock koreaboo is used to discuss the jealousy of a fictional fan when their favorite idol has entered a relationship. even though the poster has not explicitly compared themselves with another fan, even just using the mock register shows that they disapprove the jealous attitude and view themselves above it. mock koreaboo and its elitist stances constitute another order of indexicality, signaling the identity of being anti-koreaboo and what being a koreaboo represents. in this system of indexicality, korean is the nth order, being the origin of the linguistic resources that both koreaboo speech and mock korean use. koreaboo speech is the n+1st order, characterized by having korean words surrounded by otherwise non-korean utterances, typically gendered terms of address, words of love, and basic vocabulary. it is differentiated from the nth order by having this mix of language resources and by primarily being used by non-korean speakers as opposed to the korean language itself. this n+1st order indexes being a fan of k-pop and someone who knows some korean, due to its use by k-pop fans and fans of other korean media. in the eyes of other fans, koreaboo speech also indexes an obsession with korean culture and the appropriation of korean language. mock koreaboo is then the n+2nd order, innovating on koreaboo speech. this order is linguistically characterized by its use of emojis, mock forms such as oppar, and misspellings common on the internet. it is largely similar to koreaboo speech, but it is highly exaggerated and includes other mocking elements like the emojis. this indexical order incorporates the meaning of being a k-pop fan but adds the meaning of being a specific type of fan, one that disapproves of how koreaboos use the korean language and behave towards k-pop performers, due to its use by fans who actively make fun of koreaboo fans. mock koreaboo specifically indexes that they think that being a koreaboo, and therefore using koreaboo speech, is inherently wrong. through the exaggeration and mocking of these acts of appropriation, mock koreaboo clearly indexes this disapproval of the figure of the koreaboo and the linguistic practices associated with it. in addition, the elitist stances imbued into mock koreaboo help create this separate indexical order. ochs (1992) presented a view that stances are a way to indirectly index social meanings and acts. through this view, the elitist stances work to indirectly index this identity of being anti-koreaboo. while what is shown directly is the superiority over those deemed to be koreaboos, it indirectly indexes being a fan that is against the fetishization of korean culture, and the appropriation of the korean language, looking down up on the koreaboos who do these things. 6. conclusion mock koreaboo is a register among k-pop fans that utilizes elitist stances to index the disapproval of the fetishization of korean culture. through modifying a series of koreaboo speech characteristics, non-koreaboo fans are able to use pejoration to dismiss koreaboo speech and signal their elite status as compared to koreaboo fans. the use of these elitist stances with mock koreaboo create an n+2nd indexical order, innovating on koreaboo speech and indexing their disapproval of koreaboo fans. these fans are characterized as appropriating the korean language, fetishizing korean culture, and being extremely possessive of the k-pop stars that they like. mock koreaboo allows fans to distance themselves from this behavior and condemn it at the same time. this article explores how mock languages are being negotiated and used on the internet, particularly on an international scale. previously, mock languages were primarily analyzed within the context of a certain country or location, such as mock spanish and mock asian being analyzed as an american phenomenon. in this case, mock koreaboo is international, being negotiated on the internet by people from all over the world. this adds to the growing literature of how the internet can change language by allowing various linguistic resources to spread globally and be used in novel, creative ways like mock koreaboo. online language is important to study for this reason, as linguistic phenomena are used in new and exciting ways. references alim, h. samy. 2009. straight outta compton, straight aus münchen: global linguistic flows, identities, and the politics of language in a global hip hop nation. in global linguistic flows: hip hop cultures, youth identities, and the politics of language. new york: routledge. beneito-montagut, roser. 2011. ethnography goes online: towards a user-centred methodology to research interpersonal communication on the internet. qualitative research. sage publications 11(6). 716–735. https://doi.org/10.1177/1468794111413368. birney, albert & evan keogh. weeaboo. the perry bible fellowship. https://pbfcomics.com/comics/weeaboo/ (30 january, 2022). blommaert, jan. 2007. sociolinguistics and discourse analysis: orders of indexicality and polycentricity. journal of multicultural discourses 2(2). 115–130. https://doi.org/10.2167/md089.0. bonilla, yarimar & jonathan rosa. 2015. #ferguson: digital protest, hashtag ethnography, and the racial politics of social media in the united states. american ethnologist 42(1). 4–17. https://doi.org/10.1111/amet.12112. chang, woongjo & shin-eui park. 2018. the fandom of hallyu, a tribe in the digital network era: the case of army of bts. kritika kultura (32). https://doi.org/10.13185/kk2019.03213. https://journals.ateneo.edu/ojs/index.php/kk/article/view/kk2019.03213/2815 (23 april, 2021). chun, elaine w. 2004. ideologies of legitimate mockery: margaret cho’s revoicings of mock asian. pragmatics. john benjamins 14(2–3). 263–289. https://doi.org/10.1075/prag.14.23.10chu. du bois, john w. 2007. the stance triangle. in stancetaking in discourse: subjectivity, evaluation, interaction, 139–182. amsterdam/philadelphia: john benjamins. http://dubois.faculty.linguistics.ucsb.edu/dubois_2007_stance_triangle_m.pdf (12 january, 2021). ella. 2022. jinjja me는 freaking out because donghyun’s description in danjjak is an animator (correct me if i’m wrong) and choi ung is an artist and hepwpeejdbxbxbsisowpdksdbsbshjs someone help https://t.co/kineyupevw. tweet. @bomdongchan. https://web.archive.org/web/20220128215020/https://twitter.com/bomdongchan/status/14 84108006732140544 (28 january, 2022). engenie. 2021. koreaboos. let’s define it. reddit post. r/kpoprants. https://web.archive.org/web/20220124024858/https:/old.reddit.com/r/kpoprants/comment s/lsf9yu/koreaboos_lets_define_it/ (30 january, 2022). ewens, hannah. 2017. who are “weeaboos” and what does “weeb” mean? vice. https://www.vice.com/en/article/ywgxey/we-asked-j-culture-fans-to-defend-beingweeaboos (15 december, 2021). garza, joyhanna yoo. 2021. ‘where all my bad girls at?’: cosmopolitan femininity through racialised appropriations in k-pop. gender and language 15(1). 11-41-11–41. https://doi.org/10.1558/genl.18565. hallett, ronald e. & kristen barber. 2014. ethnographic research in a cyber era. journal of contemporary ethnography 43(3). 306–330. https://doi.org/10.1177/0891241613497749. herman, tamar. 2019. why k-pop is finally breaking into the u.s. mainstream. billboard. https://www.billboard.com/articles/columns/k-town/8500363/k-pop-closer-than-everamerican-pop-mainstream (23 april, 2021). hidayat, debra & z. hidayat. 2020. anime as japanese intercultural communication: a study of the weeaboo community of indonesian generation z and y. romanian journal of communication and public relations 22(3). 85–103. https://doi.org/10.21018/rjcpr.2020.3.310. hill, jane h. 1995. junk spanish, covert racism, and the (leaky) boundary between public and private spheres. pragmatics. john benjamins 5(2). 197–212. hill, jane h. 1998. language, race, and white public space. american anthropologist 100(3). 680–689. https://doi.org/10.1525/aa.1998.100.3.680. jaworski, adam & crispin thurlow. 2009. taking an elitist stance. a. jaffe, stance: sociolinguistic perspectives 195–226. jeong, sunwoo & seong-hyun yu. 2021. identity construction through gendered terms of addresses in korean. proceedings of the linguistic society of america 6(1). 829–843. https://doi.org/10.3765/plsa.v6i1.4958. lo, adrienne & jenna chi kim. 2012. linguistic competency and citizenship: contrasting portraits of multilingualism in the south korean popular media1. journal of sociolinguistics 16(2). 255–276. https://doi.org/10.1111/j.1467-9841.2012.00533.x. malik, zunera & sham haidar. 2020. online community development through social interaction — k-pop stan twitter as a community of practice. interactive learning environments. routledge 0(0). 1–19. https://doi.org/10.1080/10494820.2020.1805773. mosely, tonya & serena mcmahon. 2020. a look at k-pop’s black american influence and activism during black lives matter. https://www.wbur.org/hereandnow/2020/09/15/kpop-influence-blm (20 march, 2021). napper, arkayla. 2019. what exactly is a koreaboo and how do you know if you are one? vox atl. https://voxatl.org/what-is-a-koreaboo/ (15 december, 2021). ochs, eleanor. 1992. 14 indexing gender. rethinking context: language as an interactive phenomenon. cambridge university press 11. 335. oh, david c. 2017. black k-pop fan videos and polyculturalism. popular communication 15(4). 269–282. https://doi.org/10.1080/15405702.2017.1371309. outrageous-bottle-72. 2021. chingoose, my simjang is broken, joy unnie has betrayed me . reddit post. r/kpoopheads. www.reddit.com/r/kpoopheads/comments/p9upvd/chingoose_my_simjang_is_broken_jo y_unnie_has/ (28 january, 2022). @sgtcurrypants. 2021. all the delulu sasaengs crying : but oppar you looked at me and fell in love didn’t you?! : as i said i didn’t make eye contact with armys for 2 years now….. tweet. @sgtcurrypants. https://web.archive.org/web/20220124022444/https://twitter.com/sgtcurrypants/status/14 41506738172305408 (30 january, 2022). silverstein, michael. 2003. indexical order and the dialectics of sociolinguistic life. language & communication 23(3–4). 193–229. https://doi.org/10.1016/s0271-5309(03)00013-2. stanbts_. 2021. is my bestfriend a koreaboo? : kpoprants. r/kpoprants. https://web.archive.org/web/20220124024743/https:/old.reddit.com/r/kpoprants/comment s/lr7i23/is_my_bestfriend_a_koreaboo/ (28 january, 2022). swan, anna lee. 2018. transnational identities and feeling in fandom: place and embodiment in k-pop fan reaction videos. communication, culture & critique. oxford university press / usa 11(4). 548–565. https://doi.org/10.1093/ccc/tcy026. tracy wonwoo. 2022. : oppa, can you lend me some money? mingyu: (bank) account ohhhhhh https://t.co/j3nfbgtjb9. tweet. @tinkswonu. https://twitter.com/tinkswonu/status/1487008824636542976 (28 january, 2022). yoon, kyong. 2018. global imagination of k-pop: pop music fans’ lived experiences of cultural hybridity. popular music and society. routledge 41(4). 373–389. https://doi.org/10.1080/03007766.2017.1292819. endnotes 1 as seen in further posts: https://old.reddit.com/r/kpoopheads/comments/ovv698/i_love_my_jungkook_oppar/ 2 example: https://twitter.com/cloudreamiena/status/1439139983529369601 3 army here refers to the fandom name of bts fans. automatic opinion polarity classification of movie reviews colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by franco salvetti, stephen lewis, christoph reichenbach. automatic opinion polarity classification of movie reviews franco salvetti department of computer science, university of colorado at boulder stephen lewis department of linguistics, university of colorado at boulder christoph reichenbach department of computer science, university of colorado at boulder one approach to assessing overall opinion polarity (ovop) of reviews, a concept defined in this paper, is the use of supervised machine learning mechanisms. in this paper, the impact of lexical filtering, applied to reviews, on the accuracy of two statistical classifiers (naive bayes and markov model) with respect to ovop identification is observed. two kinds of lexical filters, one based on hypernymy as provided by wordnet (fellbaum, 1998), and one hand-crafted filter based on part-of-speech (pos) tags, are evaluated. a ranking criterion based on a function of the probability of having positive or negative polarity is introduced and verified as being capable of achieving 100% accuracy with 10% recall. movie reviews are used for training and evaluation of each statistical classifier, achieving 80% accuracy. 1. introduction the dramatic increase in use of the internet as a means of communication has been accompanied by an increase in freely available online reviews of products and services. although such reviews are a valuable resource to customers who want to make wellinformed shopping decisions, their abundance and the fact that they are mixed in terms of positive and negative overall opinion polarity are often obstacles. for instance, a customer that is already interested in a certain product may want to read some negative reviews just to pinpoint possible drawbacks, but has no interest in spending time reading positive reviews. in contrast, customers interested in watching a good movie may want to read reviews that express a positive overall opinion polarity. the overall opinion polarity of a review, with values expressed as positive or negative, can be represented through the classification that the author of a review would assign to it, if requested. such a classification is here defined as the overall opinion polarity (ovop) of a review, or simply the polarity. the process of identifying ovop of a review will be referred to as overall opinion polarity identification (ovopi). a system that is capable of labeling a review with its polarity is valuable for at least two reasons. first, it allows the reader interested exclusively in positive (or negative) reviews to save time by reducing the number of reviews to be read. second, since it is not uncommon for a review that starts with positive polarity to turn out to be negative, or vice versa, it avoids the risk of a reader erroneously discarding a review just because it first appears to have the wrong polarity. in this paper we frame a solution to ovopi based on a supervised machine learning approach. in such a framework we observe the effects of lexical filtering, applied to 1 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 reviews, on the accuracy of two statistical classifiers trained on such filtered data. we have implemented two different kinds of lexical filters, one based on hypernymy as provided by wordnet (fellbaum, 1998), and one based on part-of-speech (pos) tags. the results obtained by experiments based on movie reviews revealed that wordnet filters produce less improvement than do pos filters, and that for neither is there evidence of significantly improved performance over the system without filters, although the overall performance of our system is comparable to systems in current research, achieving an accuracy of 81%. in the domain of ovopi of reviews it is often acceptable to sacrifice recall for accuracy. here we also present a system whereby the reviews are ranked based on a function of the probability of being positive/negative. using this ranking method we achieve 100% accuracy when we accept a recall of 10%. this result is particularly interesting for applications that rely on web data, because the customer is not always interested in having all the possible reviews, but many times is interested in having just a few positive and a few negative. from this perspective accuracy is more important than recall. 2. related research research has demonstrated that there is a strong positive correlation between the presence of adjectives in a sentence and the presence of opinion (wiebe et al, 1999). hatzivassiloglou et al. combined a log-linear statistical model that examined the conjunctions between adjectives, (such as "and", "but", "or"), with a clustering algorithm that grouped the adjectives into two sets which were then labeled positive and negative (hatzivassiloglou et al, 1997). their model predicted whether adjectives carried positive or negative polarity with 82% accuracy. however, because the model was unsupervised it required an immense, 21 million word corpus to function. turney extracted n-grams based on adjectives (turney, 2002). in order to determine if an adjective had a positive/negative polarity he used altavista and its function near. he combined the number of co-occurrences of the adjective under investigation near the adjective 'excellent' and near the adjective 'poor' thinking that high occurrence near 'poor' implies negative polarity and high occurrence near 'excellent' implies positive polarity. turney achieved an average of 74% accuracy in ovopi across all domains. the performance on movie reviews, however, was especially poor at only 65.8%, indicating that ovopi for movie reviews is a more difficult task than for other product reviews. pang et al. concluded that the task of polarity classification was not the same as topic classification (pang et al, 2002). they applied naïve bayes, maximum entropy and support vector machine classification techniques to the identification of the polarity of movie reviews. they reported that the naïve bayes method returned a 77.3% accuracy using bigrams. their best results came using unigrams, calculated by the support vector machine at 82.9% accuracy. maximum entropy performed best using both unigrams and bigrams at 80.8% accuracy, and naïve bayes performed best at 81.5% using unigrams with pos tags. 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/1 doi: https://doi.org/10.25810/atv1-v819 automatic opinion polarity classification of movie reviews 3 3. statistical approaches to polarity identification there are many possible approaches to identifying the actual polarity of a document. our analysis uses statistical methods, namely supervised machine learning, to identify the likelihood of reviews having "positive" or "negative" polarity with respect to previously hand-classified training data. these methods are fairly standard and well understood; we list them below for the sake of completeness. 3.1. naïve bayes classifier the naïve bayes classifier is a well-known supervised machine learning approach. in this paper the "features" used to develop naïve bayes are referred to as "attributes" to avoid confusion with text "features." in our approach, all word/pos-tag pairs that appear in the training data are collected and used as attributes. the formula of our naïve bayes classifier is defined as where • rv is the review under consideration, • w is a word/pos-tag pair that appears in the given document, • pr(appw|class) is the probability that a word/pos-tag pair appears in a document of the given class in training data, and • bc is an estimated class. one interesting aspect of this particular application of naïve bayes is that most attributes do not appear in a test review, which means most factors in the product probability are based on what is not written in a review. this is one major difference from the markov model classifier described in the next section. 3.2. classifier based on markov models because the naïve bayes classifier defined in the previous section builds probabilistic models based on individual occurrences of words, it is provided with relatively little information regarding the phrasal structure. markov model is a widely used probabilistic model that does capture connectivity among words. this markov model classifier develops two language models: one on positive reviews and another on negative reviews. the classifier then generates two probabilities for an unseen review, one from the positive model and the other from the negative one. it compares the two probabilities and determines the classification. the following formula is the one used to compute the probability that a document could be generated using each language model. 3 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 where • rv is the review under consideration, • sn is a sentence, • 〈s〉 is the start of a sentence, • 〈/s〉 is the end of a sentence. 4. features for analysis statistical analysis depends on a sequence of tokens it uses as characteristic features of the objects it attempts to analyze; the only necessary property of these features is that it must be possible to identify whether two features are equal. the most straightforward way of dealing with the information we find within reviews would be to use individual words from the review data as tokens. however, just using the words discards semantic information about the remainder of the sentence; as such, it may be desirable to first perform some sort of semantic analysis to enrich the tokens with useful information, or even discard misleading or irrelevant information (noise), in order to increase accuracy. three basic approaches for handling this kind of data preprocessing come to mind: • leave the data as-is: each word will be represented by itself • parts-of-speech tagging: each word is enriched by a pos tag, as determined by a standard tagging technique (such as the brill tagger (brill, 1995)) • perform pos tagging and parse (using e.g. the penn treebank (marcus et al, 1994)) unfortunately, the third approach not only had severe performance issues during our early experiments, but also raises conceptional questions of how such data would be incorporated into a statistical analysis. we thus focus our analysis in this paper on postagged data (sentences consisting of words enriched with information about their parts of speech), which seems to be a good candidate for a worthwhile source of information, for the following reasons: 1. as discussed by losee (losee, 2001), information retrieval with pos-tagged data improves the quality of an analysis in many cases, 2. it is a computationally inexpensive way of increasing the amount of (potentially) relevant information, 3. it gives rise to pos-based filtering techniques for further refinement, as we discuss below. we thus make the following assumptions about our test and training data: 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/1 doi: https://doi.org/10.25810/atv1-v819 automatic opinion polarity classification of movie reviews 5 1. all words are transformed into upper case, 2. all words are stemmed, 3. all words are transformed into (word, pos) tuples by pos tagging (notation word / pos). all of these are computationally easy to achieve (with a reasonable amount of accuracy) using the brill tagger. 5. part of speech filters careful analysis of movie reviews has made it clear that even the most positive reviews have portions with negative polarity or no clear polarity at all. since the training data used here consists of complete classified reviews, the presence of parts with conflicting polarities or lack of polarity within a review presents a major obstacle for accurate ovopi. as illustration of this inconsistent polarity, the following were all taken from a single review1. "special effects are first-rate" (positive polarity) "the character is written thinly" (negative polarity) "the scenes were shot in short segments" (no clear polarity) this observation can be taken to lower levels as well. individual phrases and words vary in their contribution to opinion polarity. it may even be said that only some part of the meaning of a word contributes to opinion polarity (see wordnet filter section below). any portion that does not contribute to the ovop is noise. to reduce noise, filters were developed that use pos tags to do the following. 1. introduce custom parts of speech when the tagger does not provide desired specificity (negation and copula). 2. remove the words that are least likely to contribute to the polarity of a review (determiner, preposition, etc.) 3. reduce parts of speech that introduce unnecessary variance to pos only. it may be useful, for instance, for the classifier to record the presence of a proper noun. however, to include individual proper nouns would unnecessarily decrease the probability of finding the same n-grams in the test data. experimentation involved multiple combinations of such filter rules, yielding several separate filters. an example of a specification of pos filter rules is shown in figure 1. 1 apollo 13, a film review by mark r. leeper, copyright © 1995 mark r. leeper 5 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 (1a) copula conversion: is/* */cop (1b) negation conversion: not/* /neg (2) noun generalization: */nn /nn (3) pos tossing: */cc ø figure 1: abbreviated filter rule specification (illustrative details only) the pos filters are not designed to reduce the effects of conflicting polarity. they are only designed to reduce the effect of lack of polarity. the effects of conflicting polarity have instead been addressed by careful preparation of the training data as will be seen in the following section. 6. experiments 6.1. settings • data: taken from cornell data (pang et al, 2002) • part-of-speech tagger: brill tagger (brill, 1995) • wordnet: wordnet version 1.7.13 (fellbaum, 1998) movie reviews are used for training and evaluation of each statistical classifier. the decision to use only movie reviews for training and test data was based on the fact that ovopi of movie reviews is particularly challenging as shown by turney (turney, 2002), and therefore can be considered a good environment for testing any system designed for ovopi. the other reason for using movie reviews is the availability of large bodies of free data on the web. specifically we used the data available through cornell university from the internet movie database. the cornell data consists of 27,000 movie reviews in html form, using 35 different rating scales such as a. . . f or 1. . . 10 in addition to the common 5 star system. we divided them into two classes (positive and negative) and took 100 reviews from each class as the test set. for training sets, we first identified the reviews most likely to be positive or negative. for instance, when reviews contained letter grade ratings, only the a and f reviews were selected. this was done in an attempt to minimize the effects of conflicting polarities and to maximize the likelihood that our positive and negative labels match those that the authors would have assigned to the reviews. from these reviews, we took random samples from each class in set sizes ranging from 50 to 750 reviews (in increments of 50). these sets consisted of the reviews that remained after the test sets had been removed. this resulted in training set sizes of 100, 200, ..., 1500 (in increments of 100). html documents were converted to plain text, 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/1 doi: https://doi.org/10.25810/atv1-v819 automatic opinion polarity classification of movie reviews 7 tagged using the brill tagger, and fed into filters and classifiers. the particular combinations of filters and classifiers and their results are described in the following sections. the fact that as a training set we used data labeled by a reader and not directly by the writer poses a potential problem. we are learning a function that has to mimic the label identified by the writer, but we are using data labeled by the reader. we assume that this is an acceptable approximation because there is a strong practical relation between the label identified by the original writer and the reader. the authors themselves may not have made the polarity classifications, but we assume that language is an efficient form of communication. as such, variances between author and reader classification should be minimal. 6.2. naïve bayes according to linguistic research, adjectives alone are good indicators of subjective expressions (wiebe, 2000). therefore, determining opinion polarity by analyzing occurrences of individual adjectives in a text should be an effective method. to identify the opinion polarity of movie reviews, a naïve bayes classifier using adjectives is a promising model. the effectiveness of adjectives compared to other parts-of-speech is evaluated by applying and comparing the results on data with only adjectives against data with all parts-of-speech. the impact of at-level generalization from adjectives to synsets (or "sets of synonyms"; see "wordnet filtering", below) is also measured. the naïve bayes classifier described above was applied to: 1. tagged data 2. data containing only the adjectives 3. data containing only the synsets of the adjectives the adjectives in 3 were generalized to at-level synsets using a combination of the pos filter module and the generalization filter module. for each training data set, addone smoothing was applied to the naïve bayes classifier. table 1 shows the resulting accuracies of each data set type and size. 7 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 table 1: accuracies of naïve bayes classifier. jj means "adjectives only", wn indicates synset mapping using the generalization filter. the results indicate that at-level generalization of adjectives is not effective and that extracting only adjectives degrades the classifier. however, this does not imply that filtering does not work. adjectives constitute 7.5% of the text in the data. the accuracy achieved on such a small portion of the data indicates that a significant portion of the opinion polarity information is carried in the adjectives alone. although the resulting accuracies are better in all-pos data, adjectives can still be considered good clues of opinion polarity. 6.3. markov model three types of data are applied to the markov model classifiers described previously: 1. tagged data without any filtering, 2. tagged data with pos filters, 3. tagged data with both pos filters and generalization filters. witten-bell smoothing is applied to this classifier. 6.3.1. pos filtering one design principle of the filter rules is that they filter out parts of speech that do not contribute to the opinion polarity and keep the parts of speech that do contribute such meaning. based on analysis of movie review texts, we devised "filter rules" that take brill-tagged text as input and return less noisy, more concentrated sentences that have a combination of words and word/pos tag pairs removed from the original. a summary of the filter rules defined in this experiment is shown in table 2. 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/1 doi: https://doi.org/10.25810/atv1-v819 automatic opinion polarity classification of movie reviews 9 table 2: summary of pos filter rules wiebe et al., as well as other researchers, showed that subjectivity is especially concentrated in adjectives (wiebe et al, 1999; hatzivassiloglou, 2000; turney et al, 2003). therefore, no adjectives or their tags were removed, nor were copula verbs or negative markers. however, noisy information such as determiners, foreign words, prepositions, modal verbs, possessives, particles, interjections, etc. were removed from the text stream. other parts of speech, such as nouns and verbs, were removed but their pos-tags were retained. the output returned from the filter did not keep the original sentence structure. the concrete pos filtering rules applied in this experiment are shown in table 2. the following is an example of the sentence preprocessing: • all steve martin fans should be impressed with this wonderful new comedy • /nnp /nnp /nn be/cop /vbn wonderful/jj new/jj /nn the resulting accuracies on pos filter rules and different sizes of data sets are listed in table 3. 9 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 table 3: accuracies on pos filtering 7. wordnet filtering in non-technical written text it is uncommon to encounter repetitions of identical words; this is generally considered "bad style". as such, many authors attempt to use synonyms for words whose meanings they need often, propositions, or even generalizations. we attempted to address two of these perceived issues by identifying words with a set of likely synonyms, and by hypernymy generalization. for the implementation of these techniques, we took advantage of the wordnet (fellbaum, 1998) system, which provides the former by means of synsets for four separate classes of words (verbs, nouns, adjectives and adverbs), and the latter through hypernymy relations between synsets of the same class. 7.1. synonyms wordnet maps each of the words it supports into a synset, which is an abstract entity encompassing all words with a "reasonably" similar meaning. in the case of ambiguous words, multiple synsets may exist for a word; in these instances, we picked the first one. note that synonyms (and general wordnet processing) are only available in instances where the word under consideration falls in one of the four classes of words we outlined above. we determined the appropriate category for each word by examining the tag it was assigned by the brill tagger, not touching words which fell outside of these classes. 7.2. hypernyms for verbs and nouns, wordnet provides a hypernymy relation, which can be informally described as follows: let s1, s2 be synsets. then s1 is hypernym of s2, notation s1 � s2, if and only if anything that can be described by a word in s2 can also be described 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/1 doi: https://doi.org/10.25810/atv1-v819 automatic opinion polarity classification of movie reviews 11 by a word in s1, and s1 ≠ s2. for each of the hypernym categories, we determine a set of abstract synsets a such that, for any a ∈ a, there does not exist any s such that s � a. we say that a synset h is a level n hypernym of a synset s if and only if h � * s and one of the following holds for some a ∈ a: 1. a � n h 2. s = h and a � l s, with l < n for example, given the wordnet database, a hypernym generalization of level 4 for the nouns "movie" and "performance" will generalize both of them to one common synset which can be characterized by the word "communication." 7.3. analysis in order to determine the effects of translating words to synsets and performing hypernymization on them, we ran a series of tests which quickly determined that the effects of pure synset translation were negligible. we thus experimented with the computation of level n hypernyms with n ∈ {0 . . . 10}, separately for nouns and verbs. figure 2: hypernym generalization with 1500 reviews from each class. the x and y axis describe the level of hypernym generalization for nouns and verbs, z the accuracy we achieved. maximum concreteness at level 10 indicates no generalization. as we can see from figure 2, applying hypernym generalization to information gathered from large data sets yielded little improvement; instead, we observed a degradation in the quality of our classification caused by the loss of information. we assume that for larger data sets bigram classification is already able to make use of the more fine-grained data present. shrinking the size of our training data, however, increased the impact of wordnet simplification; for very small data sets (50 reviews and less, not shown here) we observed an improvement of 2.5% (absolute) in comparison to both full generalization and no generalization at all. increasing the size of the set of observable events by using trigram models resulted in a small gain (around 1%). interestingly, the effect of verb generalization was relatively small in comparison to noun generalization for similar hypernymy levels. 11 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 7.4. discussion our results indicate that, except for very small data sets, the use of word-net hypernymy generalization is not significantly beneficial to the classification process. we assume that this is due to at least the following reasons: • wordnet is too general for our purposes: it considers many meanings and hypernymy relations which are rarely relevant to the field of movie reviews, but which potentially take precedence over other relations which might be more appropriate here. • choosing the first synset out of the set of choices is unlikely to yield the correct result, given the lack of wordnet's specialization on our domain of focus. • for reasonably large data sets, supervised learning mechanisms gain sufficient confidence with related words to make this particular auxiliary technique less useful. considering this, the use of a domain-specific database seems to be a promising approach to improving our performance for this technique. 8. selection by ranking the probabilistic models computed by the naïve bayes classifiers were sorted by log posterior odds on positive and negative orientations for the purpose of ranking, i.e. by a "score" computed as follows: score = log pr(+|rv) log pr(-|rv) where • rv is the review under consideration, • pr(+|rv) is the probability of rv being a review of positive polarity, • pr(-|rv) analogously is the probability of the review being of negative polarity. we modified the classifier so that it: 1. sorts the reviews in the test data by log posterior odds 2. returns the first n reviews from the sorted list as positive reviews 3. returns the last n reviews from the sorted list as negative reviews the resulting accuracies and recalls on different n are summarized in table 4. 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/1 doi: https://doi.org/10.25810/atv1-v819 automatic opinion polarity classification of movie reviews 13 table 4: precisions and recalls by number of inputs the classifier was trained on the same 1500 review data set and was used with ranking on a repository of 200 reviews which were identical to the test data set. the result is very positive and indicates that adjectives provide enough sentiment to detect extremely positive or negative reviews with good accuracy. while the number of reviews returned is specified in this particular example, it is also possible to use assurance as the cutoff criterion by giving log posterior odds. 9. discussion taking all results into consideration, both the naïve bayes classifier and bigram markov model classifier performed best when trained on sufficiently large data sets without filtering. for both bigram and trigram markov models, we observed a noticeable improvement with our generalization filter when training on very small data sets; for trigram models, this improvement even extended to fairly large data sets (1500 reviews). one explanation for this result is that the filters are unable to make use of the more fine-grained information provided to them. a likely reason for this is that the ratio between the size of the set of observable events and the size of the training data set is comparatively large in both cases. however, further research and testing will be required in order to establish a more concrete understanding of the usefulness of this technique. the learning curve of classifiers with the pos filter and/or the generalization filter climbs at higher rates than those without the filters and results in lower accuracy with larger data sets. one possible explanation of the higher climbing rates is that the pos filter and the generalization filter compact the possible events in language models while respecting the underlying model by reducing vocabulary. this also explains why the plateau effect is observed with smaller data set sizes. the degraded results with filters also indicate that by removing information from training and test data, the compacted language model loses resolution. 13 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 14 10. conclusion a framework of two-phased classification mechanism is introduced and implemented with a pos filter, a generalization filter, a naïve bayes classifier and a markov model classifier. accuracies of combinations of filters and classifiers are evaluated by experiments. although the results from classifications without filters are better than those with filters, the pos filters and generalization filters are observed to still have potential to improve overall opinion polarity identification. generalization filtering using wordnet shows good accuracy for small data sets and warrants further research. using the naïve bayes classifier with ranking on adjectives has confirmed that desired precision can be achieved by dropping recalls. for the task of finding reviews of strong positive or negative polarity within a given data set, very high precision was observed for adequate recall. acknowledgements the authors would like to thank tomohiro oda for his extensive help and support during the course of all stages of the project. further acknowledgements go to larry d. blair, assad jaharria, helen johnson, jim martin, jeff rueppel and philipp wetzler for their valuable contributions. references eric brill. "transformation-based error-driven learning and natural language processing: a case study in part-of-speech tagging". computational linguistics, 21(4):543–565, 1995. tagger available from http://www.cs.jhu.edu/˜brill/rbt1 14.tar.z. christiane fellbaum. wordnet: an electronic lexical database, 1998. wordnet is available from http://www.cogsci.princeton.edu/˜wn/. vasileios hatzivassiloglou. effects of adjective orientation and gradability on sentence subjectivity, 2000. vasileios hatzivassiloglou and kathleen r. mckeown. "predicting the semantic orientation of adjectives". in philip r. cohen and wolfgang wahlster, editors, proceedings of the thirty-fifth annual meeting of the association for computational linguistics and eighth conference of the european chapter of the association for computational linguistics, pages 174–181, somerset, new jersey, 1997. association for computational linguistics. robert m. losee. "natural language processing in support of decisionmaking: phrases and part-of-speech tagging". information processing and management, 37(6):769–787, 2001. mitchell p. marcus, beatrice santorini, and mary ann marcinkiewicz. "building a large annotated corpus of english: the penn treebank". computational linguistics, 19(2):313–330, 1994. bo pang, lillian lee, and shivakumar vaithyanathan. "thumbs up? sentiment classification using machine learning techniques". in proceedings of the 2002 conference on empirical methods in natural language processing (emnlp), 2002. movie review 14 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/1 doi: https://doi.org/10.25810/atv1-v819 automatic opinion polarity classification of movie reviews 15 data available at http://www.cs.cornell.edu/people/pabo/movie-reviewdata/movie.zip. peter turney and michael littman. "measuring praise and criticism: inference of semantic orientation from association". acm transactions on information systems (tois), 21(4):315–346, 2003. peter turney. "thumbs up or thumbs down? semantic orientation applied to unsupervised classification of reviews". in proceedings of the 40th annual meeting of the association for computational linguistics (acl'02), pages 417–424, 2002. janyce wiebe, rebecca f. bruce, and thomas o'hara. "development and use of a goldstandard data set for subjectivity classifications". in proc. 37th annual meeting of the assoc. for computational linguistics (acl-99), 1999. janyce wiebe. "learning subjective adjectives from corpora". in aaai/iaai, 2000. 15 salvetti et al.: automatic opinion polarity classification of movie reviews published by cu scholar, 2004 colorado research in linguistics 6-2004 automatic opinion polarity classification of movie reviews franco salvetti stephen lewis christoph reichenbach recommended citation microsoft word paper_salvetti_lewis_reichenbach.doc making verb argument adjunct distinctions in english making verb argument adjunct distinctions in english jena d. hwang university of colorado at boulder in natural language processing, identifying a verbs argument structure is important for many tasks including parsing, text simplification, and semantic role labeling. however, while most formal theories of grammar generally agree that there is a distinction between constituents that are arguments and those that are adjuncts, linguists do not yet agree on how to define what it means to be an argument or an adjunct. nevertheless, these efforts have been fruitful in gaining a general understanding of some of the characteristics of arguments and adjuncts. this paper will explore the distinct approaches taken by the linguistics community, focusing on verb argument and adjunct distinctions. this paper explores the reasons behind why this distinction is such a difficult issue in both the semantic and syntactic community. 1 introduction most formal theories of grammar generally agree that there is a distinction that can be made between constituents that are arguments and those that are adjuncts. syntactically speaking, arguments are typically considered to be constituents that are syntactically licensed or required by the head verb of the phrase. as for adjuncts, no such restriction or requirement is necessary for them to be present in a phrase. semantically speaking, arguments are necessary participants in the event or state created by the verb and they participate in the manner specified by the verb’s subcategorization frame. adjuncts, on the other hand, unlike arguments, do not rely on the relational information conveyed by the verb. rather they comment on the general action or state of the predicating unit – the verb and its arguments. for natural language processing (nlp), identifying the verb’s argument structure is important. in nlp’s statistical parsing task, the automatic system generates the most statistically plausible parses for any given sentence. the automatic system picks from this set the most likely parse. it has been shown that providing a verb’s subcategorization information during syntactic parsing can improve performance (collins, 1999), for example by helping to resolve such issues as ppattachment ambiguity (hindle and rooth 1993; merlo and esteve ferrer 2006). in addition to the parsing task, verb argument structure information is used in distinguishing different senses of a word that are likely to be associated with a particular argument frame (dligach and palmer, 2008), in marking required or optional elements in a sentence during machine translation (deneefe and knight, 2009), and discriminating crucial information in a given document from that which is non-critical or parenthetical in text summarization and text simplification. thus, the key piece in defining verb argument structure in nlp is identifying which constituents in a given sentence should be included in or excluded from the head verb’s subcategorization frame. 1 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 consequently, drawing a distinction between what is an argument and what is not is an important task for both the syntactic resources used in the nlp communities and the lexical resources that provide human annotated data for the training and testing of the automatic systems described above. for grammar formalisms used in nlp such as tree adjoining grammar (tag) and lexical functional grammar (lfg), a clear definition of the predicate’s subcategorization frame is necessary as this is the basis for establishing the verb’s syntactic definitions (e.g. identifying tree family membership in tag and representing functional structure in lfg for a verb). furthermore, lexical and semantic resources such as framenet (ruppenhofer et al., 2010), verbnet (kipper et al., 2008), and propbank (kingsbury and palmer, 2003; palmer et al., 2005), as well as lexical resources such as comlex syntax (grishman et al., 1994) and combinatorial categorical grammar (ccg; steedman 2000) provide data for training and testing of automatic systems. effective supervised processing techniques, whether sentence parsing, machine translation or text simplification, depend highly on the quality and consistency of the annotation based on available resources. thus, this paper explores the distinct approaches taken by the linguistics community in making verb argument and adjunct distinctions and why the distinction is such a difficult issue for both the semantic and syntactic communities. additionally, this paper will briefly touch on what this difficulty means to the current the lexical resources that handle data for automatic systems. 2 semantic intuitions concerning the argument-adjunct distinction the task of making the distinction between arguments and adjuncts of a verb is in some sense a way of capturing a basic intuition that if a world event or activity must be described, such an event will necessitate participants (e.g. birthday boy in a birthday party or snow in a snowstorm) or other relevant information that is salient to the setting. and as it is in any setting, some information will be more crucial to the described event and other information will be less important (though not necessarily irrelevant). thus, linguistic intuition is that in an event described by a verb, there will be key participants without which the event would not be complete. such event will also include other peripheral information that provides descriptors of the general condition or circumstance of the state or event, which are not as central to the meaning of the verb. 2.1 thematic relationships when we invoke our intuitions of which are the “necessary” participants in a given state or event described by a verb, generally speaking, we are referencing the semantics side of the issue. take as an example a giving event as seen in the following sentence: 2 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 (1) his father gave him a computer last night for succeeding in his studies. here, we have five elements or concepts expressed in the sentence: his father, him, a computer, last night, and for succeeding in his studies. in the case of the first three elements, the relationships they have with the verb are often referred to as thematic relations. this concept of thematic relations has been well discussed by numerous studies in the linguistics literature (cf. fillmore, 1968, jackendoff, 1972, dowty, 1989). thematic relations describe the roles participants play in the event or state created by the verb. in example (1), the father is the agent or giver as he takes on the role of the giving entity, a computer would be the theme or transferred item, and him is the recipient of the computer. furthermore, these participating roles are considered required or obligatory in such a way that if they were removed from the sentence the semantics of the sentence would be incomplete, as seen in the following sentences. (2) ? his father gave. (3) ? his father gave him. for such utterances as (2) and (3) to be meaningful, the missing participant(s) (i.e. recipient and theme for example (2) and theme for example (3)) would have to be cited elsewhere and recoverable in the context. thus, these participants are considered to play a direct role in the relational information conveyed by the verb, and therefore, necessary components of the semantics of the verb. those participants that are in a thematic relationship with the verb are considered to be semantic arguments of the verb. in contrast to the arguments, the last two elements in example (1) would be considered semantic adjuncts. unlike arguments, adjuncts do not rely on the relational information conveyed by the verb. rather they comment on the general action or state of the predicating unit – the verb and its arguments. the adverbial last night and the prepositional phrase for succeeding in his studies are present because they comment on the event described by the verb and its arguments. they are there to set up the context in which the event happens regardless of the specific meaning of the verb: the adverbial sets the time in which the giving takes place and the prepositional phrase describes causal events leading up to the event, which serve as a motivation for event’s occurrence. thus, in general, the elements in the sentence that are in a thematic relationship with the verb and play a central role in the event or state presented by the verb, are considered to be arguments. these arguments are licensed and required by the verb to realize its full meaning. those elements in the sentence that do not hold a specific relationship to the verb and provide contextual information “typically information about time, location, purpose or a result of an event” (saeed, 1997) are considered adjuncts. unlike arguments, the adjuncts do not care about the specific meaning of the verb. rather, they modify the entire predicating unit. 3 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 consequently, since the adjuncts are not licensed by the verb, they are considered to be applicable to a wider range of events (c.f. cowper (1992, p.65)). that is, the same adverbial last night in (1) would retain a similar descriptive value even when used to describe other events like “scotty laughed hard at the jokes last night” or “bethany read the poetry beautifully last night”. the same could not be said about the arguments. the noun phrase a computer only serves as a transferred item in the giving event. it would take on an entirely different role specified by the verb in the context of other events (consider “the computer was damaged when it fell to the ground” where computer is a patient). as cowper (1992) and gawron (1988) note, self-evident adverbials like last night make a reference to a time and therefore are related to the verb in a specific way. however, one cannot say what thematic relation a noun phrase like the computer should hold if it is not in a relationship with a verb. 2.2 problem of semantic intuition these intuitions about the nature of argumenthood at first glance seem to be fairly straightforward. if these intuitions are indeed sufficient, then it would seem that determining the semantic representation of arguments and adjuncts should be a simple enough task. consider the following example: (4) we ate our supper on our balcony. intuitively speaking, central to the eating event are two participants: the one who eats and the entity that is eaten. the prepositional phrase provides a general location or setting in which eating takes place. it would be very simple if we could extrapolate from such an example and say that prepositional phrases like on our balcony could always be considered adjuncts as in example (4). however, as we know, this is not always the case. consider the following example: (5) i put the book on the table. the locative prepositional phrase in (4) is distinguished from the same prepositional phrase in example (5), which would generally be recognized as the argument of the verb put as it is the location in which the book is placed, an element without which the semantics of a putting event would not be complete. that is, unlike the adjunct-like prepositional phrase in example (4), on the table in (5) would have to be classified as an argument. take into consideration a few more examples, with attention to the adverbials used in the sentences: (6) scotty laughed hard at the jokes. (7) johnny hit the nail hard to drive it into the wood. (8) bethany read the poetry beautifully. beautifully. 4 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 if indeed adjuncts are characterized by the ability to be applicable to a wider range of events as the literature suggests, it would stand to reason that adverbials such as hard, quickly or beautifully should be judged adjunct-like as they maintain a similar meaning across a variety of situations. for example, hard in (6) describes the intensity by which scotty laughed at the jokes, just as it describes the intensity of force used by johnny to drive the nail into the wood in (7). however, the adverb well in (10) stands as a counterexample. (9) this child reads well. (10) this book reads well. (11) this duck eats well. here, the adverb well seems to do more than simply comment on the reading event. it has a direct effect on the interpretation of the meaning of the verb. the sentence in (10) could be used to express how gripping the said book is, which cannot be said about the sentence lacking this adverb: this book reads. in fact, the sentence is rendered nonsensical. in this particular usage of read, then, a manner adverbial such as well is a necessary component for completing the intended meaning and thus is considered to behave more like an argument than an adjunct despite our initial intuitions. moreover, in example (9), it is not clear if the adverb well is behaving like an argument as in (10) or if is an adjunct as in (8) that happens to comment on the manner in which this particular child reads. this is even more evident in example (11), where the sentence could be interpreted to mean that the animal has a good appetite (i.e. adjunct reading) or that the cooked duck is good for eating (i.e. argument reading). such a decision would likely depend on the correct identification of the reading of the sentence intended by its speaker or writer, which hopefully would be available in the context. finally, in certain cases, the distinction seems to depend on the lexical items present in the sentence. compare the following pairs of sentences in examples (12) and (13), in which each sentence carries a prepositional phrase with locative information: (12) a. i cooked the chicken on the grill. b. i cooked the chicken on the patio. (13) a. she kissed her mother on the cheek. b. she kissed her mother on the platform. (quirk et al., 1985, p.511) fillmore (1994, p.159) would consider the phrases in the (a) sentences to carry “information that fills in details of the internal structure of an event”, which he calls frame internal information, while those in the (b) sentences provide “incidental 5 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 attending circumstances of that event, the frame-external information”1. this seems to indicate that the reading of the on-phrase in the first sentences should lead to an argument-like reading. the second sentences should lead to an adjunct-like reading. the appropriate reading depends on whether the the prepositional phrase is viewed as a modifier of the location of the undergoer argument (an argument-like reading) or as a modifier of the location of the agent argument (an adjunct-like reading). 3 dimensions of distinction the level of uncertainty discussed in the previous section is reflected in the linguistic literature on how researchers think these arguments and adjuncts should be characterized. while linguists agree that there is indeed a distinction to be made between arguments and adjuncts, the researchers have not yet converged on how to define what it means to be an argument or an adjunct, and how the boundary between the two should be characterized. as przepiorkowski (1999) writes, “[a]lthough the [argument]/adjunct dichotomy is supposed to play a central role in the chomskyan version of generative linguistics, there is no generally agreed upon classification of kinds of dependents, nor is there a generally accepted analysis of adjuncts” (ibid., p.257). nevertheless, those efforts have been fruitful for gaining a general understanding of some of the characteristics of arguments and adjuncts. here are a few quotes that reflect varied definitions found in the literature: [a]rguments are something that lexical heads have; they are the central participants in the scene that the head presents, and thus in the situations the head is instantiated by. (gawron, 1988, p.111) complements tend to be (though not always) obligatory, whereas adjuncts are always optional (radford, 1988, p.263) another classic observation involving the do so anaphora is that arguments in vp are closer to the verb than other adjuncts. (culicover and jackendoff, 2005, p.128) varied as they are, the above characterizations of arguments and adjuncts display clear themes. we have already seen the first general distinction expressed by gawron (1988) earlier in this section. it speaks to our semantic understanding of arguments as central participants in the predicate. secondly, another trend in argument and adjunct distinction seen in the literature, as expressed by haegeman 1 quirk et al. (1985, p.510-511) makes a similar observation. however, he labels what fillmore calls frame-internal elements “predication adjuncts” and frame-external elements “sentence adjuncts” 6 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 (1991), radford (1988) and quirk et al. (1985), is the notion of obligatoriness and optionality, in which arguments are generally considered obligatory, while adjuncts are considered optional. a final distinction that is often found in the linguistic literature is one based on the syntactic structure of a sentence. culicover and jackendoff (2005) have chosen to use the words closer to the verb to describe the observation that arguments, in english, sit next to the verb and it is rare that other constituents are allowed to intervene between the verb and its argument. along with the notion of closeness, there is the observation that arguments form a constituent with the verb and they act within the sentence as a single unit. and as we will see in this paper, the free interchange between the terms complement and arguments, as seen in haegeman (1991) and radford (1988), also speaks to this point. 3.1 structural distinction for a discussion of the distinctions proposed by the proponents of the classical or transformational grammar, we will begin with the do so test introduced by lakoff and ross (1976) for two reasons. the do so test is a widely accepted substitution test used to distinguish arguments from adjuncts in linguistic literature appearing in many studies where the topic of argumenthood is discussed. secondly, it is a very clear and concise illustration of what role the hierarchical syntactic structure has on the distinction of argument and adjuncts. through the course of this section, we will see that for principles and parameters (p&p), minimalist theory (mp), and other transformational theories of syntax the argument and adjunct distinction is highly dependent on the structural configuration of the verb phrase. 3.1.1 the do so test the test is based on the observation that do so serves as an anaphoric substitute to the verb and its arguments (c.f. culicover and jackendoff (2005, p.124-127); cowper (1992, p.31); quirk et al. (1985, p.81-82)). the following examples illustrate the test: (14) sue cooked lunch yesterday, and fred did so today. [did so = cooked lunch] (15) *sue cooked lunch, and fred did so dinner. [did so = cooked] the test is that if do so cannot refer to the verb without one of its constituents, then that constituent must be an argument. the concept behind this test is that a verb and its arguments form a v’ and the anaphor do so must refer to the unit as a whole. for example, in (14), do so is perfectly happy to refer back to the verb and its object. however, when do so refers to the verb alone, as in (15), the sentence is rendered ungrammatical. adjuncts, on the other hand, can be included as an antecedent of the do so or they can be left out (culicover and jackendoff (2005); haegeman (1991)). this can be seen in examples (16) and (17). 7 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 (16) mary will cook the potatoes for fifteen minutes in the morning, and susan will do so for twenty minutes in the evening. [do so = cook the potatoes] (17) mary will cook the potatoes for fifteen minutes in the morning, and susan will do so in the evening. [do so = cook the potatoes for fifteen minutes] (18) mary will cook the potatoes for fifteen minutes in the morning, and susan will do so too. [do so = cook the potatoes for fifteen minutes in the morning] (19) *mary will cook the potatoes for fifteen minutes in the morning, and susan will do so the vegetables. [do so = cook] the do so test, thus, tells us the time adverbials for x minutes and in the x are adjuncts, while the object the potatoes must be an argument as the sentence in (19) fails. here are other examples in which the judgement for argument and adjunct distinction from the do so test lines up with our general semantic intuitions we have seen earlier in this paper: (20) a. *his father gave him a gift, and his mother did so a card. b. *i put a book on the table, and she did so on the counter. (21) a. we ate our supper on our balcony, and they did so on their porch. b. bethany read the poetry beautifully, but carla did so quite poorly. the test correctly predicts that the phrases in bold (20) are arguments and that the phrases in bold (21) are adjuncts. before we turn to the studies that have disputed the validity do so test, we will quickly touch on the conclusions p&p generally derives from these and other constituency tests. 3.1.2 importance of structural configuration in traditional syntactic approaches such as p&p, the do so test has often been cited as evidence for an embedded v’ or vp structure over an alternative flat structure. as culicover and jackendoff (2005) writes, syntactic behaviors based on do so substitutions seen in examples (16)-(19) lead “to the conclusion that the maximal vp consists of a nested structure of vp’s, along the lines of [(22)]:” (22) mary will cook the potatoes for fifteen minutes in the morning. (cf. (16)-(19)) 8 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 this agrees with the structures proposed by cowper (1992) and haegeman (1991). haegeman (1991) argues that a flat structure has “no internal hierarchy between constituents of v [in which] all vp-internal constituents are treated as being on equal footing” (ibid., p.79). consequently, it can only account for examples like (18), where the the do so substitutes for the entire vp. however, she argues, such a tree would only “be expected to affect either the top-node vp, i.e. the entire vp [...], or each of the vp-internal constituents, that is to say v or np or pp” (ibid., p7980). moreover, it does not provide a proper account for the fact that in examples like (16) and (17) do so only substitutes in for a portion of the vp. what would be necessary, haegman suggests, is a hierarchical structure with v’ nodes at v and at each of the pp levels, to account for substitution behaviors as seen in the do so test. in other words, the complement of the verb will be closer or local to the v (i.e. the argument will take up the complement position), the rest of the non-complements or adjuncts will join sequentially to a higher v’ node, where there is one v’ node for each of the adjuncts. this, then, begs the question of how should syntax know which constituents should occupy the which positions in the trees. the theta grid of the verb holds the necessary mapping that assigns the arguments to the complement positions (c.f. carnie, 2006, p.223-226; cowper, 1992, p.65; haegeman, 1991, p.296-297). it includes theta roles2 that the verb selects for. then through a “predictable” and deterministic association (c.f. chomsky, 1995, p.30-33; cowper, 1992, p.64-69), the theta roles are assigned to the complement positions on the tree. unlike arguments, adjuncts are not contained in the theta grid. though it is not explicitly expressed in any of the p&p literature examined in this paper, the assumption is that all other constituents are systematically attached to a v’ node which dominates, or eventually dominates depending on the number of adjuncts present, the v’ in which the verb is attached. thus, in p&p the prime distinction between the adjunct and argument is in the differences in the position in the tree occupied by the constituents, and ultimately this decision is up to the theta grid that assigns the positions (i.e. nodes sharing the same v’ as the verb). furthermore, the syntactic test above and other diagnostic tests not covered in this section3 are based on the assumption that the syntactic configuration has 2note here that the term theta role should be distinguished from the term thematic role as used in semantics. theta roles are thematic relations assigned by the verb to a particular position in the syntax (cowper, 1992). unlike semantics, in which an argument can hold more than one thematic role (e.g. the thematic role of his father in “his father gave him a gift” verb-specific label such as giver but also could be labeled with more of a general label agent). in syntax, the theta criterion indicates that there only can be one theta role assigned per constituent in a theta grid. 3aside from the do so diagnostics, there are other tests that are often used to establish argumenthood. amongst popular syntactic diagnoses are ordering restriction (radford, 1988, p.235; culicover and jackendoff, 2005, p.130; cowper, 1992, p.32), ellipsis test (radford, 1988, p.236), 9 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 much to say about the distinction. and in display of circular reasoning, these very tests have been used as evidence for the said configurational differences between arguments and adjuncts. these assumptions, as we will see in the next section are problematic. 3.2 where structural argumentation fails as a quick reminder, the do so diagnostic claims that v’ dominating the verb is an all-or-nothing unit – the anaphor do so acts on the whole v’ (e.g. (23)), and the ordering restriction states that the argument must be realized closer to the verb, before the adjunct (e.g. (24)). moreover, since do so is a substitution test, it will replace the v’ or the vp iteratively (e.g. (25)). (23) emily ate a burger on thursday, and john did so on friday. *emily ate a burger on thursday, and john did so a pizza on friday. (24) emily ate [a burger] [on thursday]. *emily ate [on thursday] [a burger]. (25) emily ate a burger slowly on thursday, ... but emily did so quickly on friday. [did so = ate a burger] ... and emily did so on friday. [did so = ate a burger slowly] ... and john did so too. [did so = ate a burger slowly on thursday] 3.2.1 culicover and jackendoff (2005) as a counter to these claims, culicover and jackendoff (2005) point out that the antecedent of do so “is not necessarily a continuous portion of another sentence” (ibid, p125), where the do so can refer to a span of constituents that are contiguous as in (26) but also noncontinguous as in (27). in addition this, they also note that the anaphor can make access to portions of the vp that could not possibly constitute a single constituent within the p&p framework (culicover and jackendoff, 2005, p.126-127). (26) robin slept for twelve hours in the bunkbed, and leslie did so on the futon. [do so = slept for twelve hours] (27) robin slept for twelve hours in the bunkbed, and leslie did so for eight hours. [do so = slept ... on the bunkbed] (28) robin broke the window with a hammer, mary did [it/the same thing/etc]4 to the table top. [do so/did it = broke ... with a hammer] and x-happen test (culicover and jackendoff, 2005, p.284-285). 4see culicover and jackendoff (2005, p.126-127) for a discussion on how anaphoras such as do it and do the same thing act like do so in standing in for a subportion of a vp. 10 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 (29) robin proved to george that mary would win the race, and bill did [it/the same thing/etc] regarding susan. [do so/did it = proved to george that ... would win the race] if indeed do so makes anaphoric reference to a complete v’ at a time, the sentences (27), (28) and (29) should not be grammatically feasible. in the case of (28), it is not only the case that the do so is accessing noncontiguous portions of the vp, but also it is leaving out the object the window, and therefore, a complement from reference. in the case of (29), the do so is substituting for far more than a single verb and its arguments – it is making reference to both arguments and adjuncts of two different verbs. what’s more, the constituent “left out” of the substitution, which by the do so test should be an adjunct, is the argument of the embedded predicate win. thus, do so cannot consistently substitute for the verb plus its arguments in a previous sentence (ibid, p.127), and it is not always the case that the adjuncts will be those that can be successfully “left out” from the substitution. in other words, the do so test fails to distinguish arguments from the adjuncts. 3.2.2 przepiorkowski (1999) a similar argument is made by przepiorkowski (1999). citing miller (1992), he argues that do so does not have to refer to the meaning of the entire maximal v’/vp node that contains the verb and its arguments. in each of the cited examples below, przepiorkowski points out that did so refers to the meaning of the verb and not the v’/vp. (30) a. john spoke to mary, and peter did so to ann. [did so = spoke] b. john spoke to mary, and peter did so with ann. [did so = spoke] c. john kicked mary, and peter did so to ann. [did so = kicked] the semantic analysis tells us that the phrase to anne in (a) is the goal of the verb speak, a verb that directly participates in the event of speaking. consequently, dropping this argument (i.e. john spoke) would alter the meaning of the utterance. on the other hand, the do so test incorrectly assesses the phrase to anne in example (30) (a) as an adjunct. a similar problem with the do so test is seen in examples (b) and (c) as well where with ann is the goal and to ann is the patient, respectively. framing miller (1992, p.96), przepiorkowski argues that “acceptability of a pp complement do so [...] is not whether or not the corresponding complement of the antecedent verb is within the vp of the antecedent, but whether or not the pp complement is acceptable as a complement for the main verb do with a thematic role compatible with that which the corresponding complement of the antecedent verb has with respect to the antecedent verb” (przepiorkowski, 1999, p.290). in other words, if the pp complement after do so is thematically the same as the pp complement in a sentence where the antecedent verb sits, then the sentence should 11 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 also be acceptable. consequently, the following sentences in which the pp complement that follows do so has a different thematic role than the one accompanying the antecedent, are judged unacceptable as seen in example (31): (31) a. *john spoke to [mary]-goal, and peter did so [for ann]-benef. b. *john kicked [mary]-patient, and peter did so [for ann]-benef. 3.3 distinction is semantic the eventual conclusion that both culicover and jackendoff (2005) and przepiorkowski (1999) arrive at is that the distinction between arguments and adjuncts belongs to the semantic, and not syntactic, level of analysis. przepiorkowski (1999) suggests that in head-driven phrase structure grammar5 all adjuncts should be treated as arguments and therefore be included in the argument structure (arg-st feature). the arg-st feature in the lexical entry of the verb specifies the arguments of the verb. through an adjunct-addition lexical rule, the lexical item would gain the adjuncts that appear in the sentence6. in a similar manner, culicover and jackendoff (2005) advocate for a flat structure for a verb, its arguments, and its adjuncts, which puts arguments and adjuncts on the same footing in their syntactic formalism (see ibid, chapter 5 and 6). it is at the conceptual structure, which contains semantic, aspectual, referential, and functional information that the distinction between arguments and modifiers of the verb is made. that is, both views argue against the structural distinction of argumenthood; rather they propose that the distinction should be made at the semantic layer of description. thus in the following section, we will return to the semantics side of the issue and look at the distinctions made based on the concepts of obligatoriness and optionality. 3.4 obligatoriness, optionality, and specificity when considering arguments and adjuncts, the general tendency is to associate obligatoriness with arguments and optionality with adjuncts. this follows from the notion, discussed in section 2.1, that argument labels are given to the participants in a verb’s event or state and adjunct labels are given to the ‘extra’ information that provide settings to the verb. however, the issue is complicated by the fact that it is not the case that all participants involved in an event or a state are always expressed in all possible sentences nor should it be that all participants involved should be expressed in all sentences. 5head-driven phrase structure grammar (hpsg) is a lexically driven constraint-based generative grammar. see pollard and sag (1994) for an introduction to the formalism. 6for the full development and use of the adjunct-addition lexical rule and other related features that feed into the addition of adjuncts to the arg-st, see chapter 9 of przepiorkowski (1999). 12 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 3.4.1 distinction amongst arguments jackendoff (2002) makes a distinction between the semantically obligatory/optional and the syntactically optional: it is often said that eat “licenses an optional argument”. however this conflates semantic and syntactic argument structure. the character being eaten is part of one’s understanding, whether it is expressed or not. this becomes clearer by comparison with the verb swallow. its syntactic behavior is identical to that of eat [...]. but although one cannot eat without eating something, one can swallow without swallowing anything. that is, swallow differs from eat in that its second semantic argument is optional. [...] more generally, we need at least to be able to say how many semantic arguments a verb licenses, and which of them are obligatorily expressed. (ibid., p.134) here, jackendoff draws a clear distinction between what is optional in the semantic sense from the syntactic sense of optionality. this discussion is further developed in culicover and jackendoff (2005, p.174-176). the distinction made is this: if a participant is implied in the semantics of the verb, then the participant is semantically obligatory for the verb. otherwise, it is semantically optional for the verb. these definitions hold independently of the syntactic expression. consider the following examples: (32) a. he swallowed/ate (the food). b. he swallowed, but he didn’t swallow anything. c. *he ate, but he didn’t eat anything. (33) a. he kicked/threw the pumpkin (down the stairs). b. he kicked the pumpkin, but it didn’t move at all. c. *he threw the pumpkin, but it didn’t move at all. according to the above definition, in example (32), since ingested item is implied in the semantics of the verb eat as tested in example (c), this participant is semantically obligatory for eat. in the same way, in example (33), the path of motion is implied in the meaning of the verb throw; it is considered a semantically obligatory argument of throw. in contrast to eat and throw, the thing swallowed for swallow and the path of motion for kick are not implied by the verbs’ meanings, and therefore these arguments are considered semantically optional for the respective verbs. whether or not the argument is syntactically expressed does not change the labels. in fact, semantically obligatory arguments that are not explicitly expressed in the syntax, culicover and jackendoff (2005) term implicit arguments (ibid., p.175). 13 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 3.4.2 distinction amongst obliques in culicover and jackendoff (2005), the terms obligatoriness and optionality refer to the distinctions made within arguments. for koenig et al. (2003), as we will see, the terms refer to the different adjuncts. these are not incompatible arguments, as we will see in the following section. rather, they represent differing cuts at the description, which are based on differing perspectives from which the issue is examined. in section 2.1, it was discussed that if a constituent is entailed by the event or state described by the verb, then that constituent is considered an argument. by this definition, the non participants would be classified as adjuncts. at first blush, this definition sounds reasonable – it generally does a good job in describing the observations we have made concerning thematic relations. however, koenig et al. (2003) brings up a problematic issue with this style of definition, namely, the classification of obliques. here is a reformulation of the definition for argument by (koenig et al., 2003, p.72): semantic obligatoriness criterion (soc): if r is an argument participant role of predicate p, then any situation that p felicitously describes includes the referent of the filler of r. thus by koenig et al. (2003)’s account, the soc is the criterion by which an argument is recognized from a set of constituents around the verb. however it is insufficient, as they note that if soc holds, then the italicized obliques in the following sentence would have to also be considered as arguments: (34) marc knits in his office during lunch. they write, that soc is too inclusive of what constituents it “lets in” as an argument. this definition as it stands would include ones that should be classified as adjuncts. if you knit, you must knit somewhere; in other words, any situation described by the predicate corresponding to the english word knit includes a location in which the event occurred. [...] again, if the soc were the sole determinant of argumenthood, the denotations of time expressions such as during lunch would qualify as semantic arguments, a conclusion contradicting most linguists’ intuitions. (ibid., p.72) the problem is that most conceivable events or states have to be located in a certain space and time (koenig et al., 2003, p.73). semantic components that describe location, time, and beneficiaries, which fillmore (1994) calls circumstantials, are entailed in just about any situation. thus, to stop the soc from overgeneralizing, koenig et al. (2003) suggest that a second definition criterion related to the specificity of the verb should be introduced: 14 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 semantic specificity criterion (ssc): if r is an argument participant role of predicate p denoted by verb v, then r is specific to v and a restricted class of verb/events. ssc forces the choice of argument to be specific to the verb. that is, ssc considers a participant role (as determined by the soc) as an argument only if the role is asked to “bear additional properties aside from those which are characteristic of the role” (ibid., p.73). take the following sentences, as example, paying special attention to the the object or the theme of the verb: (35) marc sang a song yesterday. (36) marc wrote a song yesterday. by the soc, we recognize that a song plays a role in both verbs (i.e. theme in both cases). by the ssc we recognize that a song for the specific verb sing additionally takes on a unique property that is not present in a song participating in the writing event, namely a quality that requires vocal folds. in a similar manner, the song in (36) takes on a unique property that is not present in (35), namely the written quality of the song. to say it in another way, ssc capitalizes on the fact that verbs (or certain classes of verbs) semantically ‘color’ their arguments slightly differently, and these roles that take on a different ‘color’ of meaning would be considered by ssc to be an argument. as to the adverbial yesterday, the claim is that the meaning is held constant across two or more verbs (or classes of verbs), and therefore it must be an adjunct. that is, the adverbial passes the soc, but fails the ssc. what is not discussed in koenig et al. (2003) are examples of cases in which it is either difficult to detect the “additional properities” the argument bears above and beyond the regular participant roles. here is an example in which the extra semantic coloring of the argument is not as evident: (37) marc put the apple on the porch. (38) marc ate the apple on the porch7. our semantic intuitions tells us that the oblique in (37) is an argument while the oblique in (38) is an adjunct. in order to correctly classify the locative pp in (37), according to the ssc, it would have to hold a meaning that is slightly different from that in (38). it would be up to each reader’s judgement to decide if it passes the ssc test; however, the semantic distinction for such locatives is slight at most and difficult to make. this is true not only for locatives, but also for roles such as goals, directionals, and benefactives. the question we ask here is if there is a 7intended reading: marc ate the apple while on the porch. 15 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 reliable way of testing if a given constituent holds additional semantic meaning that would classify it as an argument. in the rest of the paper, which is not discussed here, koenig et al. (2003) present linguistic judgement studies where the ssc is put to test. given the studies, it seems reasonable that such cases would have been addressed. however, they are not explicitly discussed. 3.4.3 comparing (culicover and jackendoff, 2005) and (koenig et al., 2003) in culicover and jackendoff (2005), obligatoriness and optionality were dimensions of distinctions amongst the constituents already classified as arguments. that is, obligatory arguments were distinguished from optional arguments, but they were both arguments nonetheless. here in koenig et al. (2003), they are seeking to establish a clearer definition for arguments in such a way that amongst obliques it allows as arguments only those obliques that coincide with our intuition of argumenthood. as noted above, these distinctions are not incompatible. in the general linguistic literature, there is a systematic ambiguity as to how the terms obligatory and optional are used. the first approach to defining obligatoriness or optionality is to note that amongst arguments there are those that are obligatory for the completion of the predicate’s event or state, and those that comment on the setting of said event or state. here is a graphical illustration: obligatory optional arguments/complement adjuncts semantically obligatory semantically optional this analysis is consistent with culicover and jackendoff’s (2005) view. a distinction is made within subcategorized constituents into those that are semantically obligatory and those that are semantically optional. the last box in the third row is blank as all adjuncts are considered to be optional. the second cut can be made through the syntactic core-oblique argument layer. when the cut is made through this layer, then semantic obligatoriness spans over both core and oblique arguments since either can be a potential semantic participant in the verb’s event or state. optionality spans only over the oblique arguments. here is a graphical illustration of the overlap between what is semantically obligatory/optional and what is core/oblique. obligatory optional core arguments oblique arguments the distinctions presented by koenig et al. (2003) are better represented by this analysis. in their case, the focus was on correctly separating argument obliques from adjunct obliques so that the obliques that provide setting or circumstantial information about the predicate, such as those in (34) repeated in (39), would be 16 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 classified as adjuncts. (39) marc knits in his office during lunch. thus, koenig et al. (2003) and culicover and jackendoff (2005) differ, but not incompatibly so, in their use of the terminology of obligatoriness based on the differing approaches the two studies have taken. what remains the same in both cases is that they see the distinction not as a simple dichotomy between arguments and adjuncts coinciding with obligatoriness and optionality, but rather something more complex than that. 3.5 dowty (2003) so far we have operated under the assumption that, despite the disagreement on where the distinction should be made, there is a distinction to be made. more specifically, we have examined studies that try to identify semantic and syntactic criteria by which we can say that certain constituents are arguments and certain others are adjuncts. but what would happen if we posit there is no dividing line between arguments and adjuncts? that is, even if there are decidedly “argument-y” things and “adjunct-y” things in a sentence as our intuitions seem to suggest, clear argument and adjunct categories might not exist. dowty (2003) suggests that viewing the argument-adjunct issue as a case of clear dichotomy may not be the right analysis. he argues that most distinctions that syntax and semantics have tried to draw between arguments and adjuncts have failed precisely because there is no single clean line that can be drawn between what is the argument-like constituent and the adjunct-like constituent. his solution on the problem of argumenthood is that of a dual analysis, where any given constituent of the vp can be analyzed as both argument and adjunct. as a point of illustration, dowty (2003) does a short case study on the preposition to (ibid., p.8-11). consider the following examples: (40) mary kicked the ball to the fence. (41) mary explained the memo to john. what dowty seeks to claim is that there is a semantic similarity between the two prepositional phrases headed by to: they both indicate a physical or abstract location at which something arrived as a result of the action of the predicate. however, the two phrases are also different. in example (40), the pp headed by to expresses the new location at which the object arrives at as a result of the action. this is distinguished from the to phrase in (41), whose meaning is less compositional and more argument-like than the dative phrase in (40). for example, (41) “does not mean that memo itself came to be at/near john, but only that the information contained in the memo came to be more fully understood by john, as a result of mary’s explanation” (ibid., 9). 17 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 dowty (2003) points out the usual analyses for these examples either focus on the semantic similarity or on the semantic difference between the two pps. if analysis recognizes the similarity of the phrases, the two pps are assigned with the thematic label of goal. if analysis wants to recognize that the two phrase are distinct in their semantic expressivity, however, both the argument-like or the adjunctlike uses of the preposition to is assigned a different semantic representation. the problem with these approaches, as he writes, is that the former thematic analysis fails to recognize that the two uses of to have different semantic values. the latter approach fails to recognize that the two usages are semantically related. thus, the dual analysis view, which dowty proposes, remediates both issues by positing, first, an adjunct analysis for both expressions to serve as a launching point for secondary, argument, analysis: the idea behind the dual analysis view can be thought of [...] as the claim that the locative adjunct analysis of all occurrences of to, from and other locative prepositions is a preliminary analysis which serve language-learners8 as a semantic “hint” or “crutch” to figuring out the idiosyncratic correct meaning of the complement analysis for the nonlocative instance: a preliminary adjunct analysis of the to-pp (as locative) [(40)] gives way to a complement analysis of to-pp structure as in [(41)]. (ibid., 10) dowty (2003)’s claim is that “virtually all complements have a dual analysis as adjuncts” and any adjunct can potentially be reanalyzed as a complement (ibid., 12). thus, according to his analysis, the quality of arguments and adjuncts lies in their placement along a directed continuum; every vp constituent other than the verb goes through an adjunct analysis before it can get to the argument analysis. while it’s unclear from the text when exactly it is that adjuncts would also be analyzed as arguments, at least in this way both argument and adjunct analyses should be available for any given sentence. 4 implications for nlp and conclusion in essence, all of the observations made and views proposed by the linguistics literature on the argument and adjunct distinction show that there is some degree of 8dowty (2003) has a small section devoted to the implications of the dual analysis view, and how it is cognitively a more feasible explanation for language learners who are picking up the argument and adjunct distinction. he notes that if learners first access the adjunct analysis and use it as a clue to learning the argument analysis, the learning burden would be softened. this is an interesting argument as a basis of his views, but there is only one small paragraph expanding on this claim, so there isn’t much more that could be said about it. 18 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 argument/adjunct distinction to be made, but it is seemingly without a clear solution. from a linguistic perspective, this is not problematic since it provides ample room for discussion and continuous evaluation of existing theories. however, the same cannot be said for nlp. nlp has to be able to say something definite about the argument and adjunct distinction, since machines have no intuitions to rely on. creators of resources for nlp clearly recognize the challenge. take the english penn treebank (etb; taylor 1996), as an example. etb is a syntactic resource that provides the nlp community with manually parsed corpora with phrase structure-like parses, including a wide variety of genres such as written news (e.g. wall street journal), spoken news (e.g. cnn, nbc), and weblogs. the topic of argument and adjunct distinction, as they note, is not a trivial issue. recognizing the difficulty of argument and adjunct distinction, the etb creators have chosen not to distinguish arguments and adjuncts at the syntactic level and simply make all post verbal constituents sisters to the verb: unfortunately, while it is easy to distinguish arguments and adjuncts in simple cases, it turns out to be very difficult to consistently distinguish these two categories for many verbs in actual contexts. [...] after many attempts to find a reliable test to distinguish between arguments and adjuncts, we abandoned structurally making this difference. instead we decided to label a small set of clearly distinguishable roles, building upon syntactic distinction only when the semantic intuitions were clear-cut. however, getting annotators to consistently apply even the small set of distinctions discussed here was fairly difficult. (taylor et al., 2003) their solution was to use the label closely related ‘clr’, instead, which “marks constituents that occupy some middle ground between argument and adjunct of the verb phrase. these roughly correspond to ‘predication adjuncts’, prepositional ditransitives, and some ‘phrasal verbs’ (bies et al., 1995). the definition as it stands is somewhat vague and there is no specific indication as to what predication adjuncts, prepositional ditransitives and phrasal verbs actually are. as they note, in practice, even the clr distinction was difficult to make and it was not always the case that clrs are used as consistently as perhaps etb intended. ideally, the best solution for any nlp application would be to have a specific set of features or criteria that makes a clean distinction between constituents that are arguments and those that are adjuncts. however, it is clear that there is no one set of criteria that suffices for distinctions across all verbs. if dowty is actually correct in that all elements in the sentence can be given both an argument and an adjunct analysis, it would suggest that the quest for cleanly identifying the distinction might be misguided. 19 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 references ann bies, mark ferguson, karen katz, and robert macintyre. bracketing guidelines for treebank ii style penn treebank project, 1995. url http://www.sfs.uni-tuebingen.de/˜dm/07/autumn/795.10/ ptb-annotation-guide/root.html. andrew carnie. syntax: a generative introduction. blackwell, 2006. noam chomsky. the minimalist program. the mit press, 1995. michael collins. head-driven statistical models for natural language processing. phd thesis, university of pennsylvania, 1999. elizabeth a. cowper. a concise introduction to syntactic theory. the university of chicago press, 1992. peter w. culicover and ray jackendoff. simpler syntax. oxford university press, 2005. steve deneefe and kevin knight. synchronous tree adjoining machine translation. proceedings of the conference on empirical methods in natural language processing (emnlp), 2009. dmitriy dligach and martha palmer. word sense disambiguation with automatically retrieved semantic knowledge. international journal of semantic computing (ijsc), 2(3): 365–380, 2008. david dowty. on the semantic content of the notion ‘thematic role’. properties, types and meaning, 2:69–130, 1989. david dowty. the dual analysis of adjuncts/complements in categorial grammar. mouton de gruyter, 2003. charles j. fillmore. the case for case. in emmon bach and r. harms, editors, universals in linguistic theory. holt, rinehart, and winston, new york, 1968. charles j. fillmore. under the circumstances. in proceedings of the 20th annual meeting of the berkeley linguistics society, berkeley, california, 1994. jean mark gawron. lexical representations and the semantics of complementation. new york: garland publisher, 1988. ralph grishman, catherine macleod, and adam meyers. comlex syntax: building a computational lexicon. in proceedings of the international conference on computational linguistics (coling), kyoto, japan, 1994. liliane haegeman. introduction to government and binding theory. blackwell, oxford uk and cambridge usa, 1 edition, 1991. donald hindle and mats rooth. structural ambiguity and lexical relations. computational linguistics, 19(1), march 1993. ray jackendoff. semantic interpretation in generative grammar. the mit press, cambridge, massachusetts, 1972. ray jackendoff. foundations of language: brain, meaning, grammar, evolution. oxford university press, 2002. 20 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/1 doi: https://doi.org/10.25810/g2ev-pv70 paul kingsbury and martha palmer. propbank: the next level of treebank. in proceedings of treebanks and lexical theories, växjö, sweden, 2003. karin kipper, anna korhonen, neville ryant, and martha palmer. a large-scale classification of english verbs. language resources and evaluation journal, 42(1):21–40, 2008. jean-pierre koenig, gail mauner, and breton brienvenue. arguments for adjuncts. cognition, 89:67–103, 2003. george lakoff and john r. ross. why you can’t do so into the sink. in j. mccawley, editor, notes from the linguistic underground, volume 7 of syntax and semantics. academic press, new york, 1976. paola merlo and eva esteve ferrer. the notion of argument in prepositional phrase attachment. computational linguistics, 32(3):341–378, 2006. philip h. miller. clitics and constituents in phrase structure grammar. garland, new york, 1992. martha palmer, dan guildea, and paul kingsbury. the proposition bank: an annotated corpus of semantic roles. computational linguistics, 31(1):71–105, march 2005. carl pollard and ivan a. sag. head-driven phrase structure grammar. university of chicago press, chicago, 1994. adam przepiorkowski. case assignment and the complement-adjunct dichotomy: a non-configurational constraint-based approach. phd thesis, university of tübingen, november 1999. randolph quirk, sidney greenbam, geoffrey leech, and jan svartvik. a comprehensive grammar of the english language. longman, london and new york, 1985. andrew radford. transformational grammar: a first course. cambridge university press, 1988. josef ruppenhofer, michael ellsworth, miriam r. l. petruck, christopher r. johnson, and jan scheffczyk. framenet ii: extended theory and practice. technical report, international computer science institute, 2010. john i. saeed. semantics. blackwell, 2nd edition, 1997. mark steedman. the syntactic process. the mit press, 2000. ann taylor. english penn treebank guidelines, 1996. url http://www-users. york.ac.uk/˜lang22/tb2a_guidelines.htm. ann taylor, mitchell marcus, and beatrice santorini. the penn treebank: an overview. in anne abeillé, editor, building and using parsed corpora, volume 20 of text, speech and language technology. springer, 2003. 21 hwang: making verb argument adjunct distinctions in english published by cu scholar, 2012 colorado research in linguistics 6-2012 making verb argument adjunct distinctions in english jena d. hwang recommended citation tmp.1537307026.pdf.ovzj9 functional domain of the croatian complementizer da colorado research in linguistics. june 2004. volume 17, issue 1. boulder: university of colorado. © 2004 by tamara grivičić. functional domain of the croatian complementizer da author tamara grivičić university of colorado at boulder in this paper i address the function of the complementizer da in croatian. craig (1975), frajzyngier (1995), frajzyngier (1996), vrzic� (1996), and krapova (1999) have contributed much literature about this particular complementizer. however, such studies focused on the analysis of the syntactic properties of da, that is, the structure of sentences in which this complementizer occurs. using the principle of functional transparency and independent coding means, introduced by frajzyngier and shay (m.s.), as the point of departure, the goal of the present study is to uncover the function of this complementizer and what it is the coding means of. drawing on older publications, an ample selection of sentences from the online national data corpus of the croatian language (ndc), and further elicited examples (nse) and judgments from native speakers of croatian (nsj), i will show that da is an independent coding means and that its semantic function is to code modality. in particular, i will reveal that da is a coding means of potentially realizable habitual events, and that as such it belongs to the de dicto domain. introduction according to jakab, the da-complementizer introduces a complement clause, which is frequently used in place of an infinitival complement (1999, 217).1 the complement clause therefore provides a filler, that is, additional information for the previous predicate of mention. and although da functions to introduce a complement clause, this is not its only function. but before proceeding into an investigation of the more exact function of da, a couple of important facts must first be established. a. in their work, referred to in edit jakab (1999, 219), progovac (1993) and vrzic� (1996) proposed two semantically separate but homophonous da complementizers. i will show that there is one and only one da-complementizer and provide sufficient evidence for distinction between da-comp and any other homophonous ‘da’. b. i will illustrate that although it is the case that verbs are subcategorized for specific complements, that no verb licenses the da complementizer. this in turn will lead toward the conclusion that da must be an independent marker, a morpheme that is used as a means of coding a specific semantic function of and by itself, irrespective of the verb or any other utterance it occurs with. c. once all the necessary distinctions and specifications have been made, i will introduce the testing methods, which will reveal the actual function of our complementizer. d. the last section of this paper gives a brief overview of the most crucial arguments about the functional nature of the da complementizer and offers some concluding remarks. 1. i must mention, however, that in the croatian standard s�tokavsko-jekavski dialect infinitival formations are still much preferred to da-constructions, but this is not the case with other dialects. 1 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 2 1. preliminary discussion 1.1. homophonous da(s) 1.1.1. dati ‘to give’ one ‘da,’ for which distinction must be made, is the third person singular form of the verb dati ‘to give.’ this verb is ditransitive and typically takes two np complements, which are realized as a clitic in the embedded clause and precede the verbal form ‘da.’ (1) igor hoc�e njemu dati novi auto. (nse) igor want-3sg him-dat give new-acc car-acc igor wants to give him a new car. (2) on hoc�e da mu ga da. (nse) he want-3sg da him-dat it-acc give-3sg he wants to give it to him. unlike the verbal form, da complementizer must have a clause following it and cannot occur sentence finally (see examples (15) and (18) in section 2.1 for additional comparison). 1.1.2. allegations of two homophonous but semantically distinct da-complementizers in her 1996 paper, vrzic� claims that there exist two homophonous da complementizers: “modal” and “declarative” da. she bases her statement on her observation that the modal da “is only compatible with the present tense verb…, [while] the declarative da is compatible with imperfective verbal aspect only” (307). vrzic� offers an extensive list of examples purportedly providing evidence for her claims. here are some of her examples (305-6): (3) kaz�e da vesna čita ovu knjigu. say-3sg dm vesna-nom read-3sg this-acc book-acc he says that vesna reads/is reading this book. (4) z�elim da čitam ovu knjigu. wish-1sg mm read-1sg this-acc book-acc i wish to read this book. (5) tvrdim da čitam ovu knjigu. claim-1sg dm read-1sg this-acc book-acc i claim that i’m reading this book. example (3) has a present tense matrix verb kazati ‘to say’ and is followed by a dacomplement. according to vrzić, because da follows a non-imperfective present tense 2 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 3 verb, it is declarative. in example (4), the matrix verb has the imperfective aspect, z�eljeti (imperfective) (vs. poz�eljeti (perfective)), and so she claims that the da complementizer is subjunctive-like. it is important to note here that although z�elim ‘i wish’ is coded for the imperfective aspect, that it is also in the present tense. the two categories, aspect and tense, are not mutually exclusive. vrzić’s explanation about the distribution and patterning of the da complements falls short here. example (5), shows a verb in the imperfective aspect and present tense, tvrditi (imperfective) vs. potvrditi (perfective), but the complementizer is claimed to be declarative. if vrzić were correct in her analysis, the da complementizer in the last example should be modal and so should be the mood of the sentence, yielding something along the lines of: i claim that i ought to read this book. this is not one of the possible meanings for the sentence though. let us consider some additional examples offered by jakab (219). aligning with vrzić, he claims that the presence of the declarative da brings about the indicative mood, whereas the modal da renders the modality subjunctive-like. consider example (6) below. the matrix verb is in the present tense form and the sentence in the indicative mood. the complementizer is claimed to be declarative. sentence (9) appears to be identical to sentence (6) except for the lexical entry of the matrix verb, which is also in the present tense. according to vrzić, sentence (9) should be interpreted as being in the subjunctive-like mood, in part because of the presence of a semantically different da. however, i am not convinced that either jakab’s or vrzić’s examples support their claims in full. (6) kaz�e da1 petar c�ita ovu knjigu. (jakab, 219) say-3sg da petar-nom read-3sg this-acc book-acc he says that petar is reading this book. (7) ?kaz�e da2 petar c�ita ovu knjigu. (nse/nsj) say-3sg da petar-nom read-3sg this-acc book-acc *he says that petar would read this book. (8) kaz�e da1 je petar c�itao ovu knjigu. (jakab, 219) say-3sg da aux petar-nom read-pst3sg this-acc book-acc he says that petar has read this book. vs. 3 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 4 (9) z�elim da2 petar c�ita ovu knjigu. (jakab, 219) wish-1sg da petar-nom read-3sg this-acc book-acc i wish for peter to read this book. (10) ?z�elim da1 petar c�ita ovu knjigu. (nse/nsj) wish-3sg da petar read-3sg this-acc book-acc *he wishes petar is reading this book. (11) katkad poz�elim da petar c�ita ovu knjigu. (nse/nsj) occassionally perf-wish-1sg da petar-nom read-3sg this-acc book-acc sometimes i wish for petar to read this book. *sometimes i wish petar is reading this book. if the presence of a different da complementizer yields a different mood, then substituting these complementizers should cause a change in the mood of the clause. this doesn’t appear to be the case as illustrated in the nse example in (7). placing a modal da inside of a clause doesn’t render the meaning subjunctive-like. the same applies to example (10). the declarative da doesn’t yield an indicative mood for the clause. nse solicited example (11) shows a sentence with a perfective present tense matrix clause verb and a da complement. the meaning of the sentence is judged by the native speaker to be ‘wishful thinking’ and the mood subjunctive-like with no possibility of indicative interpretation. so, at a closer look, the data presented by vrzić and jakab do not provide adequate evidence for the existence of two homophonous yet semantically distinct dacomplementizers. however, i do agree with them that sentence (6) is in the indicative mood, whereas sentence (9) codes a subjunctive-like modality. a more advanced and detailed research may ultimately prove that there exist two semantically-different yet homophonous da(s). for the time being, however, it might be more feasible to posit that the subcategorization properties of a verb, which specify verbal complement preferences, provide an environment for the appearance of a da-complement. the overall modality of a clause then depends on the combination of a complementizer with the lexical as well as apectual properties of the verb. to claim otherwise, there would have to be some kind of a formal distinction between the two sentences using ‘da1’ and ‘da2, be it syntactic, configurational or semantic difference. and there isn’t one. the above examples show that both da1 and da2 occur in identical syntactic constructions, do not change the meaning of the clauses if swapped. 1.2. verbs and their complements although it is true that some verbs prefer da-complements, i found only one case in which a verb primarily takes the da-complement. this verb is smatrati ‘to consider’ 4 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 5 (please see appendix b for examples). however, as the ndc2 examples (12) and (13) illustrate, even this verb does not exclusively take a da-complement: (12) “…i zapravo je smatram gitarom s plastic�nim z�icama.” (ndc: n146_24 15397 11) and actually it consider guitar with plastic wires …and actually i consider it to be a guitar with plastic wires. (13) “sebe uopc�e ne smatram diskriminiranom zato s�to sam z�ena.” (ndc: n135_k10 4513457 45) refl really not consider discriminated because am woman i don’t consider myself discriminated against just because i’m a woman. the fact that there is no one verb that always triggers the da-complementizer (and therefore a da-clause) contradicts the claim made by krapova and petkov (1999) that the da-comp is “licensed,” in the traditional sense of the word, by the verb that precedes it. if an item were “licensed” by another utterance, it would always have to occur in conjunction with this trigger. since no singular utterance triggers the existence of a dacomplementizer, and since its omission causes ungrammaticality (see examples (16), (19), (22), and (24) in section 2.1), da-comp must be an independent coding means. the notion of independent coding means has been proposed by frajzyngier and shay (m.s.). they assert that inflectional coding, which subsumes complementizers, is “not triggered by any other component of an utterance” (9). they say that this claim has been supported by an abundance of cross-linguistic evidence confirming that inflectional lexical items are found “in utterances lacking any potential trigger of inflection” (9). the notion of independent coding means, compounded with the principle of functional transparency, bears important implications on the function of the da complementizer. the principle of functional transparency states that “every utterance must have a transparent function within the discourse” and that this function must be transparent to the hearer (frajzyngier and shay, 3). along the lines of these two premises, the da-complementizer is an utterance that is expected to code a specific function within the discourse. it is further predicted that as an inflectional marker, da will occur independently of any trigger utterances and that its function is readily available to the native speakers. the following section shows that the above-mentioned premises hold for the da-complementizer, and it attempts to draw out its function. 2. methods in order to tease out the function of the da-complementizer, i use examples and discussion from older publications as the springboard. i further include an ample selection of sentences containing this complementizer and other relevant data (specific verbs, other complementizers, etc.), which i acquired from the national data corpus (hereafter ndc) of the croatian language, a search engine located at http://www.hnk.ffzg.hr/. i then augment the examples gathered from older publications and ndc with a set of sentences with omitted or inserted da-comps or otherwise changed 2. examples with the link style references following an example, such as ndc: n146_24 15397 11, are taken directly out of the search engine and point to the addresses from the online national data corpus (ndc) of croatian language where that sentence is located. 5 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 6 sentence structure. these i presented to several native speaker of croatian, including myself, with the goal of acquiring native speaker judgments (nsj) about the grammaticality and acceptability of the original and augmented sentences. the total of the analysis then consists of the following three steps: a. focusing on the data gathered from the ndc and older publications, i look at where da occurs and how it is used. (please see appendix c for examples) b. then using the original and augmented examples, i solicit native speaker judgements (nsj). i then incorporate these judgments, based on the native speaker intuition, and discuss any relevant findings. c. finally, i contrast sentences with da-comp against those that are identical in structure except that they contain other main croatian complementizers, namely s�to and kada. the purpose of doing this is to test the altered sentences for any change in meaning. the present premise is as follows: if a clause remains fully comprehensible and grammatical after the da-comp has been omitted (or inserted), retaining its original meaning, then this will provide evidence that da has no semantic function within the discourse utterance. if, however, the meaning changes drastically or a clause is rendered incomprehensible or ungrammatical, then da must play an important function. if the latter claim is realized, i will seek to draw out the function of da on the bases of the change in the meaning (especially focusing on the changes created when complementizers are interchanged). 2.1. instances where da occurs first and foremost, as examples (15) and (18) illustrate, the da-complementizer cannot stand alone and must always be followed by a clause. second, omission of the da– complementizer either renders the sentence ungrammatical or it may bring about a new meaning (see examples (16), (19), (22), (24)). (14) uputili su se u s�etnju da se umire. (nse) went-3pl aux-3pl refl in walk da refl calm-3pl they went for a walk (in order) to calm themselves. (15) *uputili su se u s�etnju da ø. (nse/nsj) went-3pl aux-3pl refl in walk da they went for a walk (in order) [to calm themselves]. (16) *uputili su se u s�etnju ø se umire. (nse/nsj) went-3pl aux-3pl refl in walk refl calm-3pl (no alternative meaning rendition) 6 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 7 (17) otivam tam (za) da kupja xljab. otis�ao je tamo da kupi kruh. (nsj translated penchev, 351) left-3sgm aux-3sg there da buy-3sg bread-acc he went there (so that/in order) to buy bread. (18) *otis�ao je tamo da ø. (nse/nsj) left-3sgm aux-3sg there da he went there (so that/in order) [to buy bread]. (19) *otis�ao je ø kupi kruh. (nse/nsj) left-3sgm aux-3sg buy-3sg bread-acc (alternative meaning rendition: “he left, so go buy some bread.” –not a standard dialect version) (20) kaza da dojdes�. (penchev, 348) say-3sgm da come-2sg he said that you should come. (21) rekao je da dod�es�. (nse/nsj for standard cro dialect of (20)) say-3sgm aux-3sg da come-2sg he told you to come. (22) * rekao je ø dod�es�. (nse/nsj) say-3sgm aux-3sg come-2sg (no alternative meaning rendition) (23) rekao je da c�es� doc�i. (nse/nsj) say-3sgm aux-3sg da fut come-2sg he said that you would come. (24) * rekao je ø c�es� doc�i. (nse/nsj) say-3sgm aux-3sg fut come-2sg (alternative meaning rendition: “he said: ‘will you come?’ ”) when used appropriately, da appears in several different constructions. one such construction involves adverbial clauses of purpose. within these clauses da is used to express the reasoning for the action of the matrix clause and thus provides a logical reason for the conclusion (see examples (14) and (17)). da also denotes potentially realizable events as seen in examples (21) and (23). da further appears in the function of a conjunction between two finite clauses (examples (25) 7 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 8 and (26)), but the meaning of a clause changes depending on the presence of other utterances in the embedded clause: (25) zabranih mu da dod�e ovdje. (nsj translated penchev, 348: zabranix mu da idva tuk) forbid-1sg him-dat da come-3sg here i forbade him to come here. (26) zabranih mu da bi dos�ao ovdje. (nsj translated penchev, 348: zabranix mu, za da idva tuk) forbid-1sg him-dat da would-2/3sg come-sgm here i forbade him so that he would come here. another point that ought to be made is that da-clauses are sequentially dependent. in a sentence with multiple embedded clauses, each da-clause expresses something more about the matrix predicate in its superordinate clause: (27) osjec�ala se inhibiranom[kada je trebalo[da govori o sopstvenim knjigama]] feel-1sgf refl inhibited-f when aux-3sg needed da speak-3sg about own books she felt inhibited about discussing her own books. (bibovic� (1984), 377) (28) “..ako se zna [da su tamo ubacivali bombe], [da su prijetili], [da su svec�enike if its known da aux-3pl there drop bombs da aux-3pl threaten da aux-pl priests fizic�ki smaltretirali]…” (ndc: gk9634_23 2976 6) physically abused …if it is known that they were dropping bombs there, that they were threatening, that they physically abused the priests… here we see evidence that complementizers function as connective tissue between clausal complements and that which appears before them. this is why sentences such as (29) are ungrammatical, precisely because we don’t know to which sentence da-clause is a complement.3 (29) *rekao je [da je umoran] [svima je jasno] (nse/nsj) say-3sgm aux-3sg da be-3sg tired-sgm everyone be-3sg clear ?he said [that he is tired] and this became obvious to everyone. or ? he said that it became obvious to everyone that he is tired. instances where a da-clause occurs sentence initially, or has moved and no longer immediately follows that which it is a complement of, faithfully maintains its 3. see frajzyngier (1996, 94) for additional discussion. 8 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 9 complement function to the modifying segment. subordinate clauses without a complementizer cannot move freely out of their complement positions. if moved, as example (33) shows, they render the sentence meaningless. (30) [da si mi ti to rekao], ne bih ti vjerovala. (nse) da aux-2sg me-dat you-nom that-acc say-3sgm not would-1sg you-dat believe-1sgf if you had told me that, i wouldn’t have believed you. (31) ne bih ti vjerovala [da si mi ti to rekao]. (nse) not would-1sg you-dat believe-1sgf da aux-2sg me-dat you-nom that-acc say3sgm i wouldn’t have believed you if you had told me that. (32) ti si mi to rekao. (nse) you-nom aux-2sg me-dat that-acc say-3sgm you told me that. (33) *ti si mi to rekao ne bih ti vjerovala. (nse/nsj) you-nom aux-2sg me-dat that-acc say-3sgm not would-1sg you-dat believe-1sgf *you told me this i wouldn’t have believed you. 2.2. omissions native speaker judgments show a pattern that if da is omitted, not only must the whole complement clause be removed, but also much of the superordinate sentence. and then the original meaning has been completely lost. the fact that native speakers cannot omit da and retain the intended meaning answers the question about whether da has a function at all, and whether this function is transparent to the hearer. it must be that it does, and that it is, otherwise its absence would not make any difference with respect to the meaning of a sentence. so, what is the function of da? to answer this question, let us return to some rudimentary properties of da. as mentioned in the earlier sections, a group of verbs (like smatrati ‘to consider and htjeti ‘to want’) take da-complements with a higher degree of frequency. this suggests that some verbs are subcategorized for or prefer, but do not license, da-complements. when da-comp is used in conjunction with these verbs, it behaves as a complementizer information filler. that is, da announces additional information contained within the embedded clause about its immediately preceding predicate. but this is not unexpected. other complementizers, s�to in (34) and kada in (35) for instance (what and when, respectively), also provide additional information about the matrix predicate. they do so by virtue of introducing a complement clause whose function, as frajzyngier (1996) explains, is to behave as “anaphora referring to something that was previously mentioned in speech” (100). 9 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 10 (34) rekao sam im ono [s�to mi je palo na pamet]… (ndc: jergovic_sar 97087 3) say-1sgm aux-1sg them-dat that-acc s�to me-dat aux-3sg fall on mind i told them that which first came to me… (35) …svi su osjetili olaks�anje [kada sam otis�ao]… (ndc: vj981216s 8517 61) all aux-3pl feel-pl relief-acc kada aux-1sg leave-1sgm …everyone felt a relief when i left… on the basis of what has just been illustrated, it cannot be the case that the sole function of the da complementizer is to signal the following complement clause. the presence of da, however, shows a peculiar pattern with respect to the coding between the subject of the matrix and the subject of the embedded clause. in a significant amount of data from a variety of sources (ndc, nse, and older publications), the dacomplementizer is not used when the subject of the embedded clause is coreferential with the subject of the embedded clause. (see examples (38), (39), (40), (42), (44) and (46)) (36) znas� da to kos�ta… (ndc: me980826_c02 10916 33) know-2sg da that-acc cost you know that this is expensive. (37) ??(ja) z�elim da (ja) idem. (craig 149 with nsj) i want-1sg da (i) go-1sg i want (for me) to go.’ (38) z�elim ic�i. (craig 149) want-1sg go i want to go. (39) poc�ela sam zarad�ivati prije deset godina. (craig 149) begin-1sgf aux-1sg earn-inf before ten years-gen i began to earn money ten years ago. (40) poc�ela sam zarad�ivati novce. (modified (39) to fit standard cro dialect) begin-1sgf aux-1sg earn-inf money-acc i began to earn money. (41) ?poc�ela sam da zarad�ujem novce. (craig’s 149 and nsj judgment: very uncommon) begin-1sgf aux-1sg da earn-1sg money-acc i began to earn money. 10 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 11 (42) z�elim govoriti s tobom. (craig 149) want-1sg speak-inf with you-inst i want to talk with you. (43) ?z�elim da govorim s tobom. (nse/nsj) want-1sg da speak-1sg with you-inst i want to talk with you. (44) moz�emo ic�i zajedno. (craig 150) can-1pl go-inf together we can go together. (45) ?moz�emo da idemo zajedno. (nse/nsj) can-1pl da go-1pl together we can go together. (46) moram otic�i zubaru. (craig 150) have to-1sg go-inf dentist-dat i have to go to the dentist. (47) ?moram da idem zubaru. (nse/nsj) have to-1sg da go-1sg dentist-dat i have to go to the dentist. when the matrix subject is not coreferential with the embedded subject, da must be inserted: (48) z�elim da ti ides�. (craig 154) want-1sg da you-nom go-2sg i want you to go. (49) treba da mi pitamo iskusnije ljude. (craig 151) need da we-nom ask-1pl experienced-acc people-acc we must ask more experienced people. (50) trebamo pitati iskusnije ljude. (nse/nsj; modified (49) to fit standard cro dialect) need-1pl ask-inf experienced-acc people-acc we must ask more experienced people. (51) *treba da pitati iskusnije ljude. (nse/nsj) need da ask-inf experienced-acc people-acc 11 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 12 and when matrix object is coreferential with the subject of the embedded complement, da-complement is again inserted: (52) kis�a nas je sprijec�ila da odemo. (craig 157) rain us-acc aux-3sg prevent-f da go-1pl the rain kept us from going. (53) podsjetio me je da ne zaboravim. (craig 158) reminded-3sgm me-acc aux-3sg da not foreget-1sg he reminded me not to forget. (54) dozvolili su nam da uradimo. (craig 158) permit-3pl aux-3pl us-dat da do-1pl they allowed us to do that. (55) zamolio sam ga da ostane. (craig 158) ask-1sgm aux-1sg him-acc da stay-3sg i asked him to stay. on the basis of the above examples, it is tempting to make a claim that da-comp, in addition to introducing a complement clause, functions to code switch-reference between the subject of the matrix and the subject of the embedded clause. however, to do this would be presumptuous because of the multiple instances in which such constructions are acceptable (examples (37), (41), (43), (45), (47)) even though native speakers judged them to be uncommon or not sounding quite right, but nevertheless acceptable. therefore, the difference in the meaning between these marginally-acceptable sentences and their preferred counterparts ((38), (40), (42), (44) and (46) respectively), which have no da complementizer, can not be triggered by switch reference or coreference restrictions. rather, the difference in the meaning arises because of the presence of an overt complementizer which renders the interpretation of the proposition as either factual and realizable or nonfactual and hypothetical.4 i provide support for this claim in the next section. 2.3. comparisons with other complementizers in this section, drawing on bibovic� (1984, 369), i compare da-complements with complements introduced with the other two main complementizers in croatian, s�to and kada (what and when, respectively). these three complementizers have been claimed to play an important role in the meaning of the sentences. specifically, they have been said to pattern with verbs depending on the verb’s realis/irrealis membership. amongst others, vrzic� (1996), craig (1975), bibovic� (1984), penchev (1982), have made arguments for the realis/irrealis subcategorization of verbs in croatian--each “licensing” a certain complementizer. realis sentences are theorized to be introduced 4. see frajzyngier (1995, 482) for additional supporting material. 12 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 13 with a resultative (epistemic) complementizer s�to, which belongs to the de re domain (the domain used for things that actually exit or happen); it is further said that s�to introduces finite clauses (factual and causal; clauses of reason). the function of s�to has been claimed indisputable (that is, it has been firmly and clearly established) and also holds for the present data. (56) jovan nec�e doc�i [zato s�to je bolestan]. (bibovic� 1984, 370) john won’t-3sg come because s�to aux-3sg sick-m john won’t come because he’s ill. the adverbial complementizer ‘kada’ functions to specify the temporal reference of the time of the event or discourse, (57) sibila je bila velikodus�na [kada mu je rekla istinu]. (bibovic� 1984, 375) sibila aux-3sgf generous when him-dat aux-3sg say-3sgf truth-acc sibila was generous when she told him the truth. while the da-complementizer is said to introduce irrealis predicates (that is, the nonfactive or counterfactual predication about future): (58) nesposoban je [da shvati takvu ljubav]. (bibovic� 1984, 370) incapable-3sgm aux-3sg da understand-3sg such-acc love-acc he is incapable of understanding such love. the main question that one must ask now is what exactly distinguishes the realis/irrealis quality of the matrix verbs in the above-listed examples. as examples below affirm, croatian does not subcategorize matrix verbs into realis and/or irrealis. for example, the verb čuti ‘to hear’ is used to attest a direct perception of a proposition in (61). the coding of an overt direct object in the matrix clause places čuti into the realis category, coding the de re or factual status of a proposition. but čuti also occurs with a da complement (and absence of matrix coding) in (62), and the overall matrix predicate is rendered irrealis. similarly, znati ‘to know,’ thought of as an inherently realis verb (coding direct perception), and željeti ‘to wish,’ an irrealis verb, both pattern with the da complementizer. therefore, it can not be the inherent realis/irrealis property of a verb that renders the meaning of a sentence hypothetical, but rather the presence of the da complementizer. examples (59) through (63) are native-speaker solicited examples, and have been molded after frajzyngier and shay examples testing epistemic coding with verbs of perception (m.s. 152-155). (59) znam da ga nije bilo. know-1sg da him-acc not be-pst i know he was gone. 13 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 14 (60) z�elim da ga nema. wish-1sg da him-acc not i wish he were gone. (61) c�ula sam ga kada je govorio u tvornici. matrix coding of direct perception hear-1sgf aux-1sg him-acc when spoke-3sgm in factory-loc i heard him when he was speaking in the factory. (62) c�ula sam da je govorio u tvornici. no matrix coding (indirect perception) hear-1sgf aux-1sg da spoke-3sgm in factory-loc i heard that he spoke in the factory. (63) **c�ula sam ga da je govorio u tvornici. matrix coding and da-comp hear-1sgf aux-1sg him-acc da spoke-3sgm in factory-loc **i heard him that he was speaking in the factory. the above sentences illustrate that the absence of matrix coding and the presence of the complementizer da code indirect perception in croatian. absence of coding of the embedded clause subject in the matrix clause indicates that the embedded clause is in the hypothetical mood. matrix coding is used to signal realis modality and since verbs in croatian are not inherently realis/irrealis, matrix coding alone indicates the realis mood, whereas the absence codes irrealis. furthermore, as expected, example (63) illustrates that simultaneous matrix coding (realis) and da-complementizer (irrealis) cannot occur together. in the following sentences with s�to-complement and da-complement substitutions, the function of da-comp becomes evident. (64) sretan sam s�to te vidim. (nse/nsj) happy-m aux-1sgm s�to you-acc see-1sg i am glad to see you. i am happy because i see you now (65) sretan sam da te vidim. (nse/nsj) happy-m aux-1sgm da you-acc see-1sg i am glad to see you. i am happy whenever i get an opportunity to see you. (if i see you) (66) tes�ko mu je s�to z�ivi sam. (nse/nsj) difficult him-dat aux-3sg s�to live alone-m it is hard for him to live alone. it is hard for him as a consequence of living alone. 14 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 15 (67) tes�ko mu je da z�ivi sam. (bibovic� 1976, 9) difficult him-dat aux-3sg da live alone-m it is hard for him to live alone. it is hard for him to live alone. (68) kao s�to sam ti rekao, mladic�u… (ndc: stiks_dvorac 11568) as s�to aux-1sg you-dat tell-m young man-voc as i told you, young man… (69) kao da sam ti rekao, mladic�u… (nse/nsj) as da aux-1sg you-dat tell-m young man-voc as if i had told you, young man… matrix verbs in examples (64), (66), and (68) are followed by a s�to complement. the same matrix verbs are then followed by a da complement in (65), (67), and (69). these pairs of sentence (one with the s�to and the other with the da complementizer) are identical but for the complementizer. it is precisely this singular distinctive feature that brings about the change in the sentence meaning. therefore, it must be that the s�tocomplementizer encodes the matrix clause as factual or epistemic, while the dacomplementizer functions as a habitual, indirect or hypothetical de dicto marker.5 if this analysis is correct then we should expect that a de re complement (i.e. s�to-complement) cannot be conjoined with another resultative clause, but de dicto can. and this is the case as the following sentences indicate: (70) tes�ko mu je da z�ivi sam i zato z�ivi sa sinom. (nse/nsj) difficult him-dat be-3sg da live-3sg alone-m and therefore live-3sg with son-inst it is hard for him to live alone and that is why he lives with his son. (71) *tes�ko mu je s�to z�ivi sam i zato z�ivi sa sinom. (nse/nsj) difficult him-dat be-3sg s�to live-3sg alone-m and therefore live-3sg with son-inst * it is hard for him as a consequence of living alone and that is why he lives with his son. 3. conclusion: toward a common function da-complementizer introduces a complement clause which says more about the immediately preceding predicate. therefore, each time a da-clause is used it is a comment on something else. the da complement cannot be interpreted in isolation (*’da je umoran’ ?that he is tired…, unlike ‘umoran je’ he’s tired), and is therefore said to be pragmatically dependant on the matrix clause. regardless in what construction da is used, its combinatorial possibilities are limited to occurring with certain verbs that are 5. de dicto and de re domains are explained in more depth in frajzyngier (m.s., 210-11), frajzyngier and jasperson (1991), and frajzyngier (1991). 15 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 16 subcategorized for da complements. as we have seen, this does not mean that a verb licenses a specific complementizer since no one verb takes only one kind of a complement. however, when the da complementizer follows a matrix verb, it functions to code the matrix clause for its modal property – hypothetical mood. this function of coding deontic modality is most clearly evident when the da complementizer is juxtaposed with the epistemic s�to complementizer. in otherwise identical sentences, the deontic modality is brought forth via da complementizer, and s�to complementizer renders the epistemic status of proposition. whether there also exists a declarative da, in addition to the modal da, is a matter that needs to be researched further before any conclusive claims may be made. at this time it is unclear whether the indicative reading is merely a consequence of an alternative da complementizer, or whether perhaps the verb’s lexical and aspectual properties outweigh the modal properties of the complementizer. the data presented in this paper however provide sound evidence that the dacomplementizer belongs to the deontic (hypothetical) de dicto domain (a domain of speech that can extend to hypothetical), whose specific function is to signal potentially realizable habitual events.6 6. see darden (1997, 90) for additional discussion. 16 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 17 appendix a abbreviations 1 first person 2 second person 3 third person acc accusative case aux auxiliary dat dative case dm declarative mood f feminine fut future tense inf infinitive inst instrumental case loc locative case m masculine mm modal mood ndc national data corpus of croatian language nom nominative case nse native speaker example nsj native speaker judgment perf perfective aspect pl plural refl reflexive sg singular voc vocative case 17 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 18 appendix b 30m corpus of contemporary croatian language (test version) 20.11.2002 05:57:37 corpus: 30m_test results of query: smatram click the source name to see the wider context ---------------------------------------------------------------------------> varalo tobože značajne političke osobe, zbog čega i danas smatram da je posao destabilizacije hsls-a bio znatno širi n159_08 12424 1 judskog tijela, osobno sam se opredijelila za fitness jer smatram da je zdraviji, ženstveniji i primjereniji nama že n154_22 6849 2 jeravanja našeg pastorala prema odgoju odraslih vjernika, smatram prioritetnim usredotočivanje naših snaga na stručn gk9712_i 4117 3 ="../slike/blueball.gif" width=12 height=12> prije svega, smatram da je sedam bolničkih kreveta za urološke bolesti me971217_v05 12541 4 d kulturne povijesti bosne i hercegovine". i inače osobno smatram prijateljstvo s lovrenovićem svojom najvećom povla n153_09 24029 5 zdavač vrhova svjetske književnosti
budući da smatram da članak u vašem listu "izdavački skandal u škols n135_r04 9984 6 alo treće, i time će, vjeruju, oprati svoju savjest.
smatram da se bez osnivanja stalnog međunarodnog suda neće n131_14 21099 7 z obzira što grad zagreb ni od koga ništa ne traži. stoga smatram da je dosta tog licemjerja prema gradu zagrebu.

smatram da je još preuranjeno razgovarati o bregovićevu go n143_19 29676 9 vu, iako se i tu jako pretjeruje. ja, na primjer, i dalje smatram, kao što sam i prije smatrao, da hrvat nije naobra n149_10 19453 10 ojem se može sve svirati, mekša je od gitare i zapravo je smatram gitarom s plastičnim žicama.
extra: recit n146_24 15397 11
ja sam uvijek za umjerenost, pa tako i u izjavama. smatram da je svaka godina za hrvatsku teška dok se ne rij vj981227t 16739 12 ševila. nikada se ne bih odlučila za tako nešto, jer sebe smatram plesačem. osim toga, imam određeni otpor spram nje n129_k10 14998 13 ga mu čestitam, što sam već mnogo puta učinio osobno, jer smatram kako je on tu ispao mnogo veća žrtva od mene.
n138_23 15825 14 je i skinuti odgovornost i sa sebe, ali bit ću iskren jer smatram da se oko te emisije diže nepotrebna prašina. dok n160_17 9384 15 štetama za duševne boli mogla bi se napisati disertacija: smatram da bi ih se trebalo svesti na simboličnu jednu kun n160_04 21983 16 obrim književnim djelom popularnost se ne postiže. osobno smatram da je jedan od odličnih hrvatskih pisaca pavao pav vj981102k 17307 17 ku scenu u nas? mislite li da ima medijske slobode?
smatram, što se tiče tiskovina, da su medijske slobode čak vj990105t 13732 18 sti broj nastavnih tjedana kod pojedinih nastavnika.
smatram da je bila nemoralna i protuzakonita odluka g. pug vj990102i 4347 19 ah, već i zbog geografske blizine tog područja hrvatskoj. smatram da je problem herceg bosne nemoguće riješiti prije n147_12 18399 20 ojim te se svijete ii". ovaj film, bez obzira na nagrade, smatram svojim najvrednijim ostvarenjem", priznaje andrea. n137_22 15927 21 rodnu nagradu ondas za radijsku emisiju "zvižduća priča". smatram to značajnijim od emisije "ljudi smo". osvajanje t n145_25 16607 22 je nezahvalna zvijer. hrane joj treba sve više i više.

smatram da istinu ne bi trebalo uokvirivati, jer to stvara gk9647_20 5923 23 dvadeset i pet godina koliko radim u svijetu menagemanta, smatram se uspješnom poslovnom ženom.
• kakav je vaš i vj981202z 3919 24 lištu.
do pokrivanja troškova ekipe nije došlo, što smatram izuzetnom štetom. riječki teatar je umjetnički rel vj981215k 9865 25 no na koncert gorana bregovića ne namjeravam otići jer ga smatram antiglazbenikom —
kada je davor št n159_20 17494 26 je od jučer'. to znači da veoma poštujem prošlost u modi, smatram je velikom inspiracijom svih modnih kreatora na sv n152_19 5566 27 jivati ili navijati za ovu ili onu struju.
osobno ne smatram da je u takvom hdz-u moguće rasplesti krizu na dem n150_01 15843 28 am reći jednu veliku istinu, možda će se neki iznenaditi. smatram da je jedan od najvećih hrvatskih ljudi, intelektu vj981231t 5020 29 eg se u hrvatskoj dodirnete čezne za razvojem. dakako, ne smatram kako je hrvatska totalitarna država, već društvo n n139_01 15341 30 e ući u koaliciju s vidom bogdanovićem i ostalom oporbom. smatram da se nitko neće usuditi sam vladati u dubrovniku n145_05 13231 31 teru. »ne, to nije to. mrzim se klanjati na pozornici. to smatram najsramotnijom tradicijom na svijetu. nije potrebn somen 56737 32 spunjena očekivanja i opravdane prigovore organizatorima, smatram da interliber ima smisla i da mu mi nakladnici svo vj981114k 7754 33 određenog pravnog isustva, kao i povjesničarskog znanja, smatram da mogu procijeniti kakvi su dokazi koje imam prot n138_06 6472 34 se nikada ovdje nije ni pojavio. i što to najamnik želi? smatram da je papa dovoljno mudar i dovoljno mlad da sam, n159_09 19163 35 boraviti. rekao mi je zabrinuto kako ima loš predosjećaj. smatram da su najkvalitetniji hrvatski igrači pružili manj n148_16 17705 36 anak i dobro su im poznate teme o kojima smo razgovarali. smatram potrebnim još jednom naglasiti da izaslanstvo u či n146_r04 17633 37 upina teška, čekaju na teške utakmice s jugoslavijom, ali smatram da te utakmice treba shvatiti samo kao dvoboje koj vj981220s 13867 38 o je došlo "svježeg zraka" je drugo pitanje i trebalo bi, smatram, problematizirati što mislimo pod "svježim zrakom" me981223_v06 1929 39 an film, napokon održan je, eto, nekoliko godina kasnije. smatram da je za sve potrebno najmanje dvoje da bi se nešt n154_20 12419 40 e, pa i policiju i redovite organe gonjenja. kako sebe ne smatram tajkunom, nisam shvatio da se primjedba gospodina n154_11 22441 41 2">u svezi s aktualnim raspravama o pravosuđu u hrvatskoj smatram potrebnim istaknuti da bi eventualno ustrojavanje vj981203i 3732 42 ćom povlasticom u zadnjih deset godina, i u tom se smislu smatram čistim ratnim profiterom. — politiku smatram otrovom i loše je to što su političari postali ono n155_14 12809 49 mislim da je najteže živjeti u laži i prikrivanju istine. smatram da smo i mi ljudi kao i svi ostali. upravo zbog to n141_16 15025 50 irate li osnovati udrugu homoseksualaca?

da, jer smatram da je vrijeme da i ti ljudi budu prihvaćeni od zaj n141_16 19890 51 e 4-5 u opasnosti je postignuti ugled hrvatskog nogometa. smatram da je mikša dobar predsjednik hns-a, i ne vidim ra n153_12 20753 52 18 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 19 appendix c 30m corpus of contemporary croatian language (test version) 01.10.2002 19:06:56 corpus: 30m_test results of query: da click the source name to see the wider context ---------------------------------------------------------------------------> obodnijem prijevodu, protekcija...â«

⻞eliš kazati da se saulus-paulus, time, i po drugi put zaplotnjački oko f91lausic0 82406 1 le, a zrinski se să˘m sebi učini kao da je od stakla, kao da mu se udovi uleđeno pokreću, ohladnjeli do smrtne obamr f92stahuljak 505487 2 an zagreba već su ušle u službenu proceduru. ekštajn kaže da u razgovorima s ministrima matešom i šeparovićem nije s vj981227p 19342 3 skoči u pomoć u akciji suzbijanja ovih štetnika, pomislim da će i nas početi noću žderati, poput scena u filmovima s me980819_v01 2074 4 je odgovornosti predsjednika«.
mnogi će se prisjetiti da su u rujnu 1993. godine šestorica hrvatskih intelektual vj981120o 7515 5 određena psihoza, ako se zna da su tamo ubacivali bombe, da su prijetili, da su svećenike fizički maltretirali, da gk9634_23 2976 6 toga sjeđenja spremaju za nju. u svakom slučaju očito je da su jedno u drugo zaljubljeni, što se vidi i po gorljivi delorko 110843 7 anu. mi, dakako, tu situaciju želimo promijeniti. ne zato da bismo i mi došli u situaciju da imamo prednosti kao i c vj981113s 12200 8 , a 1990. za podsekretara za državnu sigurnost. zato kaže da ga, doduše, nije proganjao bivši režim, ali da 'nije bi vj9811171 15351 9 dovoljan i sretan, ali nastavi nešto mrmljati, tek toliko da ne iznevjeri svoju prirodu: -ohladit će se ručak... brucic 5175 10 ) te desetak bolesničkih kolica. danas bi ih, zbog sumnje da kreveti nisu tamo gdje im je mjesto (?), želio prebroji me980819_v02 3451 11 as je i do sada u niz navrata prevario svojim prijetnjama da će napustiti izbornički posao. ali što će biti s hrvats n136_01 18679 12 natog vaterpolista vjeke kobeščaka, optužila je ivu radić da je pobijedila zahvaljujući namještanju rezultata. praši n145_24 4795 13 znih stranačkih činovnika. a znao sam kroz dosta indicija da je predsjednik donio tu odluku. on je čak jednom od vis n153_01 31185 14 anstvom hrvati u bih stimuliraju se da napuštaju bih tako da prijeti opasnost da sami sebe pobijede, da ostanu bez s n141_k09 7374 15 rihvatili demokratska pravila: zato su i dobili mogućnost da postave svog premijera


nije pljaš f92stahuljak 563644 19 a« dosta toga ponavlja. ipak, organizatoru treba priznati da, za razliku od ugostitelja, svaki njihov godišnji skup vj981102o 12750 20 ticu. zadovoljstvo mu se prelije licem.

rekao sam da dobro gađaš, samo ti konj ne valja reče frankopan i p f92stahuljak 187719 21 jeme tražio pravu osobu. tako su prošla već tri tjedna, a da odgovor nije stigao u županiju. no i kad stigne, hoćemo me960612_v04 2608 22 onda mi on pozivno mahne rukom i meni se odjednom objavi da ja, samo li se priberem, umijem i mogu po vodi hodati, f91lausic0 52112 23 -šapućem sebi, a vidim jasno da se ta svinjska glava ne da: ta, u meni tako jasno živi to sneno oko, oko prazno i segedin 352459 24 omažu. cijela moja obitelj zahvaljuje."

bog nas poziva da u tijeku došašća budemo blagovjesnici ubogima, da iscje gk9650_k01 4738 25 arstvo, naša književnost sve što nam je dalo puno pravo da i u minulim stoljećima, kao i sada, zahtijevamo da bude gotovac 230600 26 u glasova, ili se radi o nečem drugom? više je nego čudno da ni račan ni gotovac, niti drugi stranački čelnici, nisu n130_05 11702 27 u srijedu, s jednim glasom protiv takve odluke. to znači da bi oko 150 tisuća umirovljenika kojima je obustavljena vj981217p 17401 28 ve reći... ali kad sam ti već ovo rekel... bilo je i toga da sam imal drogu za prodaju. znal sam si uzeti 10-15 gram me980826_c02 9247 29 dstavnicima međunarodne zajednice neprestano objašnjavati da se bez kažnjavanja glavnih zločinaca i svih zločinaca, n138_k09 9292 30 i poviješću, pokušavamo, i na to imamo pravo, učiniti sve da, potpuno ravnopravno, bez ikakve majorizacije ili bilo n151_09 21690 31 a i sam obred vjenčanja bude jednostavan i dostojanstven, da se u nj razborito unose drevni svadbeni običaji različi gk9623_07 2323 32 e/blueball.gif" width=12 height=12> kakvu terapiju?! znaš da to košta, a toliko novaca mama i ja nemamo. a da si hod me980826_c02 10916 33 osrće. šute. nitko ne plače. ni djeca. samo pate. osjećam da me peku tabani. krvare mi stopala. bole i žare, ali pod stojsav_dnev 81197 34 a gumeni čep na stolu. pustite čeličnu kuglicu kroz cijev da padne na čep i odskoči. izračunajte stupanj korisnosti fizika 22390 35 ne strane, među nizom vila na obroncima jurjevske, gotovo da zrači svojom mondrijanskom strukturom pročelja vila mei maroevic_zg 142538 36 pture. na tom putu, u toj radnji ima toliko raznolikosti, da se konfiguracija mora »čitati« kao da je misao ili misa dragojevic_c 22051 37 su ti dječaci i djevojčice nekoliko sati bili zatvoreni i da se u toj navali možda iskazala stiješnjenost, stegnuti dragojevic_c 100738 38 ogim indoeuropskim jezicima) znači bilo kakav prijem ili, da kažemo, koktel -na malo višoj razini. u australiji, m nick_diploma 83348 39 adore čiji engleski često nije bio bolji od njegova, tako da nije imao komplekse prema njima. na večere koje je rije nick_diploma 191103 40 nama ne škodi. njegov je sadržaj u pozitivnome značenju: da se čini dobro, a ne samo da se izbjegava zlo. očito, po pozaic_cuvar 148581 41 telefonirati njoj, reći kako je uljanicu konačno upalio i da je sve kako treba, i da joj se ništa nije dogodilo. odu rehak_preobr 55672 42 ljudima dan kako bi mogli prikriti svoje misli; tvrdio je da je richelieu varao, ali nikada nije lagao, dok metterni nick_diploma 258213 43 gnuća -prema tome i transplantacija, trebala ići za tim da čovjeku omoguće ljudskije, čovjeka dostojnije postojanj pozaic_cuvar 208545 44 pa da ga je neka granata, neki geler usmrtio, čini mi se da bih lakše prihvatio --, nastavi marijan, -ali ovako.. rehak_preobr 184827 45 iz društva otići tada kada je najugodnije. nije potrebno da odlazak gostiju izgleda kao stampedo, dopušten je i čak nick_diploma 367966 46 ljučuje prva dva sadržaja i dodaje novi: ubojstvo čovjeka da bi ga se konačno oslobodilo od svake boli, patnje i nei pozaic_cuvar 287357 47 h razaranja, otišao je u luku kako bi definitivno utvrdio da je jedrilica izgorjela. i zaista, na mjestu gdje je bil rehak_preobr 125936 48 u svom obmanjivanju i pokazivanju nestida ide tako daleko da svoje zalaganje za jugoslaviju '90. uspoređuje s tuđman me970430_m01 9120 53 oja nemaju potporu u realnoj sferi, gospodarstvu, ukazati da nešto može biti drugačije. stvari imaju svoju dinamiku vj981227t 9188 54 19 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 20 i jednog 20-godišnjaka.
srpski mediji izvijestili su da su u subotu poslije 13 sati u selu dašincu, pokraj deča vj9901111 5863 55 mu se organiziraju takozvani razredni ispiti ako zaključi da su profesori dosadni, a program nepotrebno slušati. nar gk9625_58 4888 56 je amerikance pa sam ga čak dva puta pokušala nagovoriti da se preseli u sad jer sam sigurna da bio i tamo napravio n134_02 29125 57 o društveno pozitivnoj vrijednosti te se zalagao da ga se ne shvaća kao suvremeni društveni fenomen koji bi gk9714_f01 2655 58 pak, stavovi većine istraživača podudaraju se u shvaćanju da se tehnika odnosi na sredstva za rad, znanja proizvodnj mraovic 101815 59 elini napravi nužan i radikalan zaokret? dojam je, naime, da je oporba prilično mlako reagirala na aferu oko dubrova n134_03 12691 60 ebe svrstava izvan europe. čak je i albanija smogla snage da njihova vlada donese odluku da za školstvo iz bdp-a izd n135_03 11532 61 ije toliko pogodilo i toliko povrijedilo koliko činjenica da su oni koji su željeli uništiti svaki hrvatski trag na gk9710_28 561 62 romjenjljivosti republičkih granica, on bi tražio od mene da se u zaključak unese malo drukčija formulacija. umjesto n148_20 15289 63 om i washingtonskom sporazumu. tako bi izbjegli situacije da međunarodna zajednica umjesto nas mora donositi zakone n155_05 9400 64 nabroj>b) ilovaste pjeskulje -budući da se ovo tlo sastoji od pijeska koji je pomiješan s nešto eko_polj 144555 65 imovini koju su oni dugo godina stvarali. 12. zalagati se da se svakom umirovljeniku individualno povećava mirovina stranke 899625 66 javnosti. tako su, htjeli to ili ne, naprosto prisiljeni da se sučele s vinovnicima te afere, a, po svemu sudeći, i n130_07 12829 67 ranaka da posjete washington. u javnosti je stvoren dojam da bi taj odlazak oporbenjaka u ameriku hdz mogao iskorist n133_03 22989 68 nost i rad na području matematičke znanosti.

nadamo se da će naše pisanje i pisanje akademika babića biti doprino gk9630_57 3544 69 pobjeđujem! pobjeđujem i ovo »in medias res«. vi mislite da se varam! ne, ne, moj slatki gosparu, ja bih i danas bi segedin 380576 70 vnoga zbora, slovenskoga parlamenta, što je bilo osnovano da bi se ispitale i ustanovile sve okolnosti i posljedice vj981106p 15542 71 o snimio fotoreporter jednoga tamošnjeg tjednika. čini se da je i ovoga puta zakazala vlast, jer je ona (a ne netko vj981202g 19096 72 al: kako komentirate vojni dio sporazuma?
mislim da su i dosadašnji vojni sporazumi davali rezultata, prije n155_05 18746 73 i u vremenu kad je u svijetu umjetniku bilo poželjno reći da je iz sarajeva nikad to niste istaknuli?
to mi n149_17 14214 74 m priznavanju hrvatske, premda je meni itekako bilo jasno da je to toliko dugo čekano priznanje, u stvari »debelo« k vj981209t 8717 75 dručju. također bi bio promašaj kad bi koji čovjek mislio da vanjskim svojim držanjem, poštivanjem drevnih običaja, gk9625_f02 4500 76 e i u duhovnoj obnovi društva.

uvjeti za rad

da bi se netko mogao služiti metodom hagioterapije, potreb gk9627_f01 1532 77 sredne priprave na svetu godina 2000. stoga bi bilo dobro da se na prvu nedjelju došašća u svim našim crkvama navije gk9648_22 9507 78 i vrhovnom sudu, a vrhovni sud ovaj put nije donio odluku da obrvan nema pravo žalbe nego da je njegova žalba neosno n132_07 8850 79 istri su ocijenili kao neutemeljen, jednako kao i tvrdnju da njemačka izvozi više na tržište eu-a nego ostale članic vj981209t 14196 80 ržavali od kritike dok se eu stvarao, i to sve u interesu da ideja europskog ujedinjenja uspije. međutim, sada, nako vj990115g 10577 81 vnodušno prema životinjama, a da je papa nedavno izjavio " da kršćani moraju promijeniti svoj odnos prema životinjama gk9652_k02 4864 82 i, koji su u jutarnjim izdanjima dnevnih novina pročitali da su brena i boba u zagrebu, čekali već od 10 sati ujutro n134_05 5934 83 nskom dvojbom: komu od dvojice svojih najbližih suradnika da vjeruje u aferi oko dubrovačke banke, hrvoju šariniću i n128_07 14239 84 ništvu sfrj prevladao optimizam, pa se razmišljalo i tome da se tita na daljnji oporavak prebaci u njegovu rezidenci n133_04 21454 85 za osmišljavanje suvremene »duhovne igre«, već i težnjom da se suptilnost modernih književnih postupaka približi ko zmegac_b 228001 86 era.
s obzirom da ga nije bilo na programu, znali smo da će prvi dodatak biti obavezan valcer »na lijepom plavom vj990102k 5231 87 log za to šarinićevo odsustvo iz zagreba ili pak procjena da nakon političkog debakla u hercegovini ne može potpisat n132_08 5299 88 ina sršan.

više detalja podupire pretpostavku da je netko požar podmetnuo. u štaglju nije bilo električn me980311_c03 1543 89 e da je iskoristite na najbolji mogući način. zato tvrdim da u životu nije odlučujući talent, nego rad i volja. najv n135_05 17671 90 kog spota u programima hrvatske televizije u jednom danu. da hrvatska televizija nije pretjerala, uvjerili smo se i n146_r03 9505 91 f" width=12 height=12> trebat će ići posuđivati okolo. ma da sam to mogla znati i posljednju kunu bih dala za osigur me980311_c03 3190 92 d jedne do pet godina zatvora. kazneni zakon, naime, kaže da je za to kazneno djelo, odavanje poslovne tajne, predvi n154_08 6200 93 e. no, da bi se ovaj proces ispravno odvijao, potrebno je da preživači imaju dovoljno »voluminozne« hrane (trava, si eko_polj 631635 94 ta, a odredbe traže do to bude između 20 i 30 kandidata i da pri sastavljanju te liste župnik vodi računa o tome da gk9710_45 3657 95 duhovnom i u materijalnom pogledu, civilizacijski, radno, da sve što o hrvatskoj govorimo i sanjamo bude na novom pu gotovac 286576 96 aljka napravimo mali otvor i usmjerimo nakupinu elektrona da proleti kroz valjak. slučaj a: u rezonantnoj šupljini n radioter 146419 97 jačanje imunosti pridonijeti uništenju tumora, tek ostaje da se utvrdi. s druge strane, poticanje na proliferaciju t radioter 1133235 98 drug marko mu ovaj put ne dopusti.

'pričekaj malo, da ipak tu stanemo', poče on i onda upita marijana: 'rekao f90ivin0 134955 99 edan problem muči predsjednika tuđmana. on, naime, smatra da granić i valentić ne podupiru dovoljno iskreno njegovu n129_10 21570 100 mir šoljić pokušao u den haagu nagovoriti zatočene hrvate da javno podupru hdz bih na rujanskim izborima i ograde se n146_r04 11372 101 tv-a, stidljivo progovorilo tek iza prognoze vremena. kao da ksenija urličić s time nema baš nikakve veze i da su za n153_04 14326 102 je u ponedeljak čelnik kluba hdz-a vladimir šeks dodavši da su razgovori s dr. tuđmanom zasad samo u sferi špekulac vj9812221 12197 103 amjenika, a u ono doba nije ni ispitivao, jer se smatralo da za te poslove još nije spreman. uglavnom je odlazio u i f90ivin0 67545 104 nog trga sa željom, zapravo potrebom da bar malo prošeće, da i stvarno doživi prestanak zatvorske skučenosti. usput f90ivin0 198911 105 to »zovete«, ali već je bilo kasno: riječ je odletjela... da, a ipak sam bio svjestan kako postoji razlog zašto sam segedin 28648 106 20 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 functional domain of the croatian complementizer da 21 references bibovic�, ljiljana. 1984. ‘the structural possibilities of serbo-croatian related to the english structure-adjective + prepositional sentential compliment.’ linguistica 24, 369-382. bibovic�, ljiljana. 1976. ‘the infinitive as subject in english and serbo-croatian.’ yugoslav-serbo-croatian-english-contrastive-project-reports 10, 3-19. craig, colette. 1975. ‘on serbo-croatian complement sentences.’ yugoslav-serbocroatian-english-contrastive-project-studies 6, 148-164. croatian national corpus. url: http://www.hnk.ffzg.hr/cnc.htm, december 2002. darden, bill j. 1997. ‘on the prehistory of the slavic nonindicative.’ balkanistica 10, 8194. frajzyngier, zygmunt. 1996. grammaticalization of the complex sentences: a case study in chadic. amsterdam: benjamins. frajzyngier, zygmunt. 1995. ‘a functional theory of complementizers.’ in joan bybee and suzanne fleischman eds., modality in grammar and discourse. amsterdam: benjamins, 475-502. frajzyngier, zygmunt. 1991. ‘the de dicto domain in language.’ approaches to grammaticalization. ed. by elizabeth c. traugott and bernd heine. volume i amsterdam & philadelphia: benjamins, 219-251. frajzyngier, zygmunt and robert jasperson. 1991. ‘that clauses and other complements.’ lingua.83.133-153. frajzyngier, zygmunt and erin shay (m.s.). systems interaction in language. jakab, edit. 1999. ‘is pro really necessary? a minimalist approach to infinitival and subjunctive(-like) constructions in serbo-croatian and hungarian.’ in dziwirek, coats &vakareliyska, eds., formal approaches to slavic linguistics: the seattle meeting 1998, 205-224. ann arbor: michigan slavic publications. krapova, i. and v. petkov. 1999. ‘subjunctive complements, null subjects, and case checking in bulgarian.’ in dziwirek, coats, & vakareliyska, eds., formal approaches to slavic linguistics: the seattle meeting 1998, 265-287. ann arbor: michigan slavic publications. maldzhieve, vyara. 1990. ‘characterization of the da-construction in bulgarian regarding its functional equivalents in slavic languages.’ contrastive-linguistics vol. 14, no. 4-5, 213-217. mihaljevic�, milan. 1997. ‘yes-no questions in croatian church slavonic.’ suvremena lingvisitka 23, 1-2(43-44), 191-209. nazor, anica. 1973. ‘slavic syntax. selected works from slavic studies.’ slovo 23, 222225. penchev, iordan. 1982. ‘the conjunctions da and za da ‘in order to’ in standard bulgarian.’ international journal of slavic linguistics and poetics: studies for edward stankiewicz on his 60th birthday 17 november 1980, 347-353. progovac, ljiljana. 1993. ‘locality of subjunctive-like complements in serbo-croatian.’ journal of slavic linguistics 1, 116-144. 21 grivic?ic?: functional domain of the croatian complementizer da published by cu scholar, 2004 colorado research in linguistics, volume 17 (2004) 22 vrzic�, zvjezdana. 1996. ‘categorical status of the serbo-croatian “modal” da.’ in j. toman, ed., annual workshop on formal approaches to slavic linguistics: the college park meeting 1994, 291-312. ann arbor: michigan slavic publications. 22 colorado research in linguistics, vol. 17 [2004] https://scholar.colorado.edu/cril/vol17/iss1/5 doi: https://doi.org/10.25810/2e02-q527 colorado research in linguistics 6-2004 functional domain of the croatian complementizer da tamara grivičić recommended citation microsoft word paper_grivicic.doc approaches to "context" within conversation analysis colorado research in linguistics. june 2009. vol. 22. boulder: university of colorado. © 2009 by [author full names as they appear in the author line]. approaches to "context" within conversation analysis joshua raclaw university of colorado this paper examines the use of "context" as both a participant’s and an analyst’s resource with conversation analytic (ca) research. the discussion focuses on the production and definition of context within two branches of ca, "traditional ca" and "institutional ca". the discussion argues against a single, monolithic understanding of "context" as the term is often used within the ca literature, instead highlighting the various ways that the term is used and understood by analysts working across the different branches of ca. the paper ultimately calls for further reflexive discussions of analytic practice among analysts, similar to those seen in other areas of sociocultural linguistic research. 1. introduction the concept of context has been a critical one within sociocultural linguistics. the varied approaches to the study of language and social interaction – linguistic, anthropological, sociological, and otherwise – each entail the particulars for how the analyst defines the context in which language is produced. goodwin and duranti (1992) note the import of the term within the field of pragmatics (citing morris 1938; carnap 1942; bar-hillel 1954; gazdar 1979; ochs 1979; levinson 1983; and leech 1983), anthropological and ethnographic studies of language use (citing malinowski 1923, 1934; jakobson 1960; gumperz and hymes 1972; hymes 1972, 1974; and bauman and sherzer 1974), and quantitative and variationist sociolinguistics (citing labov 1966, 1972a, and 1972b). 1 to this list we can add a number of frameworks for doing socially-oriented discourse analysis, including conversation analysis (ca), critical discourse analysis (cda), and discursive psychology (dp). these last three frameworks served as the focus for a critical dialogue on the nature of context in socially-oriented discourse analysis (billig 1999a, 1999b; schegloff 1997, 1998, 1999a, 1999b; wetherell 1998). as the papers by billig, schegloff, and wetherell exemplify, the conversation analytic understanding of context is often framed as contentious (and a "methodological limitation") outside of scholarship in the ca tradition. such criticisms have emerged within pragmatics (e.g. searle 1986), linguistic anthropology (e.g. blommaert 2001, 2006; briggs 1997; bucholtz 2003), sociology (e.g. cicourel 1981; lynch 1985), 1 ervin-tripp (1996) further unpacks the varied approaches to "context" seen in sociolinguistics. 1 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 2 and a range of scholarship in other traditions of socially-oriented discourse analysis. these critical discussions highlight a widespread understanding of ca's view of context as monolithic. in this paper i argue that it is not, however, enough to refer simply to a "conversation analytic" approach to context. rather, it is necessary to refer instead to the numerous approaches to context seen across the different branches of ca that have emerged over the past two decades. 2 the issue of context in ca is compounded by the fact that there are many varied aspects of an interaction that analysts understand as being elements of its context: the sequential organization of a singular utterance, the social milieu of the larger interaction (e.g. its institutional setting), and the membership categories and identities ascribed moment-by-moment to participants and others are all potentially relevant to the interactants (and thus, to the analyst). though allowing for the relevance of each of these elements to an interaction, the majority of work within "traditional" sequential ca has focused primarily on only the first, the organization of an utterance in relation to the elements of talk occurring immediately prior. however, other branches of ca – those conducting analyses of interactions within both institutional and cross-cultural settings, work that examines the interactional aspects of social organization in children's peer groups, and conversation analytic research informed and motivated by a feminist politic – have adopted an analytic focus that also demonstrates the relevance of such contextual elements as cultural practices and epistemology, or sociological categories like gender or race. common among each of these branches of conversation analytic research is the analyst's understanding of each as a local, demonstrably relevant participant's resource rather than an a priori construct. it is not only the scope of the contextual elements that varies across these branches of ca, however. analysts working in branches other than "traditional" ca often frame both their data and findings as sensitive to ethnographic issues and/or to discipline-specific epistemologies (such as the knowledge that dominant gender and sexual identities operate hegemonically, and are thus not oriented to by participants in quite the same way as other sociological categories and identities). as i have argued elsewhere (raclaw 2010), traditional ca may also 2 scholarship over the past two decades has increasingly made space for variations in analytic practice and scope within ca, with two named varieties in particular emerging as distinct enterprises: "institutional ca" and "feminist ca". i use the not-entirely satisfactory term "traditional ca" to refer to work in sequential ca that is more often referred to simply as "ca" (given its broad association with the framework as a whole), of the kind that heritage (1997) contrasts with institutional ca. this branch of ca has elsewhere been referred to as "core," "classic," or even "schegloffian" ca (e.g. blommaert 2001; mcilvenny 2002; speer 2001). to these "branches" i also consider what i term "cross-cultural ca" and "peer organizational ca", though work within these two areas is not generally organized as distinct, named varieties of ca (though see moerman 1988). while not discussed within this paper, it is also worth noting the existence of what might be termed a sixth branch of ca, conversation analytic work falling under the purview of interactional linguistics (e.g. ochs, schegloff & thompson 1996). i delineate the five fields under discussion here largely to highlight the differences in scope, and in the reflexive treatment of context, seen across the larger field of ca. 2 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 3 make use of what we might alternately refer to as "ethnographic," "member," or "common-sense" knowledge as an analyst's resource, though this practice has yet to receive the same reflexive discussion it has in other branches of ca. 3 the use of these forms of knowledge is highly visible in institutional, cross-cultural, and peer organizational ca in large part because they emerge from settings outside of the analyst's own; while the use of discipline-specific epistemologies is highly visible in feminist ca due to the influence of feminist theory and methodology. it may be that the similar forms of knowledge deployed in traditional ca is far less marked a resource simply because it stems from a mundane source, and because the move to reflexivity seen elsewhere in sociology (even within em) has yet to reach conversation analysis as a whole. one aim of this paper is to work towards clarifying exactly what is meant by the concept of "context" within conversation analytic research; in particular, how analyst approaches to context may vary in some ways, while remaining constant in others, across the different branches of ca outlined above. a secondary goal is to contribute to the dialogue concerning those ways that context is analyzed and invoked – and how the analyst's understanding of context is informed – across these separate branches of conversation analytic research. due to the relatively limited space allotted to this article, the discussion that follows is limited to the two largest branches of conversation analysis: traditional ca and institutional ca. the present discussion is not the first to suggest that context may be approached differently by the analyst working within these different branches of ca. hammersley (2003, p. 774) notes that "the question of the role of ‘ethnographic context’ has arisen in a particularly sharp form in debates about the study of ‘institutional talk’ and about the relationship between ca and feminism. on the first, see boden and zimmerman (1991), drew and heritage (1992), hak (1995) and psathas (1995); on the second, see edley (2001), kitzinger (2000), speer (1999, 2001a, 2001b) and stokoe and smithson (2001)." other work, such as maynard (2003) and arminen (2005), devote entire chapters to discussing the necessity for institutional ca to engage with ethnographic practices normally seen as beyond the purview of traditional ca. 3 two points of terminological clarification: what is understood in the ca and em literature as constituting "common sense knowledge" (for example, knowledge of membership categories and their predicates) is not always clearly analogous with other forms of what has been termed "member knowledge," and i do not wish to conflate these terms. however, as smithson and stokoe (2001) have noted, the forms of natively gleaned member knowledge drawn upon by conversation analysts are often referred to by either term within the literature without clear distinctions made between them. additionally, in keeping with the use of the term "ethnographic knowledge" by conversation analysts such as maynard (2003) and schegloff (1987, 1992, 2006) to refer to this same arena of knowledge, i do not wish to conflate ethnographic knowledge with ethnographic methods (e.g. fieldwork, participant observation). rather, i hope to point out the similarities in the epistemological foundations of work done in institutional, cross-cultural, and peer organizational ca (which may well involve ethnographic fieldwork) and traditional ca (which almost universally does not). 3 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 4 my discussion will begin with a look at the foundations for the development of ca-specific understandings of context within traditional ca, which provides a basis for many of the methodological similarities seen among other branches. 2. foundational research: ca and the ethnomethodological project one starting point for these foundations is in the sociological framework of ethnomethodology (em) developed by harold garfinkel, which had a profound influence on the early work of harvey sacks. 4 while ca and em comprise separate research programs that have increasingly diverged since sacks' passing in 1975, the ethnomethodological influence on the development of ca (heritage 1984), and on the contemporary "social theory" of ca (heritage 2008), is well cited. hammersley (2003) frames one particular aspect of ca's approach to context as one of its "basic methodological commitments … a refusal to attribute to particular categories of actor distinctive, substantive psychosocial features – ones that are relatively stable across time and/or social context – as a basis for explaining their behaviour," and traces this commitment back to an influence of the ethnomethodological framework. hammersley sees garfinkel's insistence that em employ rigorous, scientifically-sound analytic methods as being a clear influence on ca's understanding of context as a locally established rather than static or a priori aspect of an interaction (seen, for example, in the ubiquitous analyst's question in ca, "why that now" (schegloff and sacks 1973, 299)). this priority to establish scientific rigor within sociology is echoed in sacks' own writings on ca methodology, particularly in how he advocates that analysts approach their data "without bringing any problems to it" and engage in the practice of "unmotivated looking" (1984). (we see this too in more recent conversation analytic practices, such as investigating talk for evidence of the "next-turn proof procedure" or the strict avoidance of "theoretical imperialism" within the analysis itself.) schegloff (1992a) also argues that this shared stance between em and ca on analytic rigor influenced conversation analytic understandings of context. however, schegloff additionally notes a divergence reflected in the explicitly "anti-positivist and anti-science" stance that garfinkel set forth for ethnomethodology, while "sacks sought to ground the undertaking in which he was engaging in the very fact of the existence of science" (p. xxxii). hammersley (2003) also argues that the phenomenological influence on ethnomethodology helped to shape conversation analytic understandings of context. this is a point taken up most clearly by arminen (2005), who notes that 4 there are, of course, other potential influences on sacks' early work that may have contributed to the development of ca's distinct understanding of context, such as the symbolic interactionism pioneered by goffman (e.g. 1959). however, a full review of these earlier influences (those external to "ca proper") is outside the scope of the present paper. 4 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 5 the phenomenological concept of "bracketing" – "where the question of what the world ‘really’ is is closed off and the inquiry instead concerns the appearance of the world and how it is constructed as it appears to us" – contributed to how "ca inquiries suspend knowledge about the external context of interaction, and study the way participants make the context relevant for themselves in the course of an ongoing interaction. in this way, ca studies the endogenous construction of context … unselective and unmotivated data exploration allows the analyst to notice features and possible phenomena without a theory-drive pre-selection of the focus" (9). 3. foundational research: sacks and membership categorization analysis similar to the distinctions made between ca and em, lines are also often drawn between the analytic frameworks of membership categorization analysis (mca) and sequential ca (despite both emerging from the writings and lectures of harvey sacks). the present discussion largely maintains this division as set forth in schegloff (1992a, 2007), lerner and kitzinger (2007), and elsewhere in the literature. however, arminen's (2005) dissenting view is also worth noting in light of its relevance to understandings of context within ca: "this strict division and the whole notion of ‘pure’ ca (as distinct from mca) is misleading and inadvisable. moreover, separating talk from its context goes against all the basic ideas of ca, according to which the context-renewing properties of talk amount to the endogenous construction of context, as parties orient to the ‘context’ through the management of talk-in-interaction as an observable part of doing social actions in the context" (5). for arminen, then, membership categories (and their corresponding devices and predicates) are (perhaps necessarily) part of the context of an interaction as experienced by the participants; to work within the framework of sequential ca while excluding the analysis of these categories is thus a contradictory effort. the tenets of mca also have much to do with how context has been approached in institutional ca, since institutional identities like doctor/patient or teacher/student can have a significant bearing on the sequential organization of interaction. the division between mca and sequential ca has much of its roots in schegloff's (1992a) concern with the "culturalist tenor" (p. xliv) of sacks' spring 1966 lectures (sacks 1992), from which the groundwork on mca first emerged, and his argument that sacks abandoned the study of "category-bound activities" in his later work because of their "promiscuous" (p. xlii) analytic use. however, sacks' understanding of membership categories as emergent aspects of the context of an interaction is still a vital one within ca, particularly in work on (primarily institutional) settings where the demonstrable relevance of these 5 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 6 categories is itself a tool for creating and establishing context. 5 for example, watson (1997) notes that speaker categories such as "caller" and "called" are made relevant throughout the course of an interaction, and that membership within one of these categories entails specific category-bound rights and obligations (e.g. called speaks first, caller is category-bound to initiate a preclosing). watson takes this argument so far as to argue that mca is truly inseparable from ca, as "categorical organization is intrinsic to...turn ordering" (p. 16), and "the procedural apparatus sacks formulated in his early work concerning mcds (membership category devices) can work to explicate the operation of turn-generated categories" (p. 30). in watson's view, then, the tenets of mca necessarily inform conversation analytic understandings of context. 4. understanding context within ca beyond sacks' solely-authored work, numerous other early writings within the ca canon contributed to the core understandings of context within the framework. of these ideas, one of the most profound and influential has been the understanding that the organization of talk is both "context-sensitive" and "context-free." these terms convey the view that a particular spate of talk is necessarily shaped by its local, immediately surrounding context, yet the practices employed within that spate of talk can be investigated across different social and interactional contexts. schegloff (1972) provides perhaps the first published description of talk-in-interaction as context-sensitive, noting that "to say that interaction is context-sensitive is to say that interactants are context-sensitive" (emphasis in original). here, as in much of his future work, schegloff argues that context is as much of a sense-making tool for participants as it is for analysts, and that ca must therefore investigate "how participants analyze context and use the product of their analysis in producing their interaction" (p. 115). the understanding that interaction also exhibits a context-free operation emerged later in sacks, schegloff, and jefferson (1974), which describes the turntaking mechanism of talk-in-interaction as both "context-free and capable of extraordinary context-sensitivity" (p. 699). the authors expand on the situatedness of talk within a locally-determined context, describing how "conversation is always 'situated' – it always comes out of, and is part of, some 5 it should be noted the "promiscuous" nature of mca has been challenged by a number of analyst's who support its use alongside sequential ca. silverman's (1998) review of mca research argues that this promiscuity and risk is not "inevitable," especially when mca is combined with conversation analytic work on sequential organization. additionally, watson (1997) argues against schegloff's claim that sacks' later work shifted away from membership categorization in favor of sequence organization, claiming that this view "shows an overly-selective attention [to] the empirical topics of sacks' work rather than its general conceptual commitments" (2). 6 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 7 real sets of circumstances of its participants." however, they also highlight the import of the context-free operation of interaction in noting how "it is undesirable to have to know or characterize such situations for particular conversations in order to investigate them" (p. 699). though a crucially important aspect of the conversation analytic framework, this "context-free" characterization of talk is also frequently misread by critics of ca, who cite it as evidence that the framework pointedly ignores the context of an interaction (especially what is often described as "macro" levels of context, e.g. the relevance of gender or the operation of power). 6 however, as arminen (2005) notes, "every subsequent conversational move renews our understanding of the prior move so that each turn both orients to a preceding context but also recreates the context anew. therefore, a purely formal context-free description of a conversation remains impossible. instead, conversation analysis amounts to discerning the participants’ intersubjective understandings of the course of conversation as it evolves moment by moment, as the participants orient themselves to the social action" (2). arminen’s discussion also argues for the critical import of context for doing ca, as we saw above. this argument leads to yet another aspect of conversation analytic understandings of context, that talk-in-interaction is both shaped by the context of an interaction, and ultimately works to (re)produce the context of the interaction. in this sense, talk is what heritage (1984) describes as "doubly contextual" in being both "context-shaped" and "context-renewing" (p. 242). heritage later expands on these descriptions by noting that "it is context-shaped because its contribution to an ongoing sequence of actions cannot be adequately understood except by reference to the context in which it participates … this contextual aspect of utterances is significant both because speakers routinely draw upon it as a resource in designing their utterance and also because, correspondingly, hearers must also draw upon the local contexts of utterances in order to make adequate sense of what is said. … communication action is also context-renewing. since every current utterance will itself form the immediate context for some next action in a sequence, it will inevitably contribute to the contextual framework in terms of which the next action will be understood. in this sense, the context of a next action is inevitably renewed with each current action. moreover each current action will, by the same token, function to renew (i.e. maintain, adjust or alter) any broader or more generally prevailing sense of context which is the object of the participants’ orientations and actions" (1989, 22-23, emphasis in original). within these early descriptions, heritage formally defines "context" both in a "micro" sense, insofar as it consists of the localized environment in which the 6 see the series of papers by billig, schegloff, and wetherell described above for a comprehensive discussion of this critique. 7 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 8 interaction occurs, as well as at a more "macro" level, what he describes as "the ‘larger’ environment of activity within which that configuration analysably occurs" (p. 22-23). in later work, heritage (1997) does away with these two levels of distinction to provide a definition of context that is far more reliant on the sequential organization of an interaction: "sequences of actions are a major part of what we mean by context, that the meaning of an action is heavily shaped by the sequence of previous actions from which it emerges, and that social context is a dynamically created thing that is expressed in and through the sequential organization of interaction" (p. 223). across all branches of ca, context is seen as encompassing those immediately local aspects of an interaction that are produced within it (rather than as an external influence to it). in this sense, context is a constantly renewable and alterable resource for participants. 5. relevance and orientation because context is a participant's resource, its relevance to the talk underway – and speaker orientations to this relevance – should be demonstrable within an analysis. schegloff has made this point clear in much of his writing, though three papers in particular (1987, 1991, 1992b) are often cited as explicitly laying out a "program" for how to analyze aspects of context within sequential ca. 7 as he notes in the first of these, it is only through close attention to these orientations and displays of relevance that the analyst determines what aspects of the interaction hold meaning for the participants themselves, and it is these aspects of the interaction that are of particular interest to ca: "this form of analysis takes seriously the relevance of the fact that the interactions we are examining were produced by the parties for one another and were designed, at least in part, by reference to a set of features of the interlocutors, the setting, and so on, that are relevant for the participants. the fact that these interactions are structured and progressively restructured by the participants’ orientations does not serve…to make "objective" analysis irrelevant or impossible; it is precisely the parties’ relevances, orientations, and thereby-informed action which it is our interest to describe…under the control of the details of the interaction in which they are realized. it is what the action, interaction, field of action are to the parties that poses our task of analysis" (3, emphasis in original). this type of demonstrable relevance is what schegloff (1991) terms "procedural relevance." schegloff notes in this paper that everything said within a particular context does not necessarily attend to identities potentially mobilized by 7 this is not to say that schegloff's other work doesn't explicitly touch upon the issue of context within ca, of course. schegloff (1988), for example, lays out some of the finer points of ca's approach throughout most of its conclusion, and more recently, schegloff (2009) does much of the same in presenting a critical review of some of two chapters from sidnell (2009). 8 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 9 that context in a procedurally relevant way. for example, the institutional identity of a police officer who has self-identified as such to a caller, while talking to them on a police line, may not be relevant to a caller solely by virtue of either their selfidentification or the institutional setting of the call. rather, the identity is made relevant through the activity of the talk itself. schegloff introduces the "paradox of proximateness" as a means for guiding analysts to procedurally relevant aspects of the talk: "if it is to be argued that some legal, organizational or social environment underlies the participants’ organizing some occasion of talk-ininteraction in some particular way, then either one can show the details in the talk which that argument allows us to notice, and which in return supply the demonstrable warrant for the claim by showing the relevant presence of the sociolegal context in the talk; or one cannot point to such detail" (64, emphasis in original). schegloff (1992b) also draws on the paradox of proximateness in comparing conversation analytic understandings of context to those understandings seen throughout much of the social sciences. in the latter case, context is often divided into what he terms "external" (or "distal") forms and "intra-interactional" (or "proximate") forms. outside of ca, such "macro-level" sociological categories as gender, sexuality, social class, race, and ethnicity are generally seen as "external" aspects of context, as are the various "institutional matrices within which interaction occurs (the legal order, economic or market order, etc.) as well as its ecological, regional, national, and cultural settings" (1992b, p. 195). all of these contextual features then "shape" the course of an interaction from outside of it. within ca, however, all relevant context can be termed "intra-interactional," given speakers' demonstrable orientation to it within the course of the interaction itself. the paradox of proximateness problematizes the above microand macro-level distinctions, as what we generally think of as "external" context must be shown to be "intra-interactionally" relevant in order to analyze it, at which point (due to being "intra-interactionally" relevant) its "external" status is moot. similarly, if a potential aspect of context cannot be shown to be "intra-interactionally" relevant, then it should be seen as "external" to the interaction (i.e. no longer worth the analyst's consideration). schegloff's discussion also reinforces the need for analysts to attend to procedural relevance, as it allows ca to "show from the details of the talk or other conduct in the materials that we are analyzing that those aspects of the scene are what the parties are oriented to" (p. 110). reiterating the idea that context is ultimately a participant's resource, schegloff (1992b) argues that for an outside party – like the analyst – to make sense of this context, they must attend only to that which participants may be shown to attend to. that participants demonstrably orient to relevant aspects of the talk can be seen in a "next-turn proof procedure" (hutchby and wooffitt 1998) by which a next turn at talk is seen to show a speaker's orientation to a prior turn as accomplishing a particular action. within a question-answer sequence, for example, that a question receives an answer is evidence of the second speaker's 9 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 10 orientation to the action of the first as being a question. within a "traditional" understanding of ca, then, speakers will similarly orient to the relevance of aspects of the talk such as its institutional character, or the relevance of gender to a particular sequence or spate of talk. however, the methods used by members to display orientation to contextual elements outside of sequential organization are not all recognizable through the next-turn proof procedure, nor has the traditional ca literature made clear what types of interactional practices signal orientations to other elements of context. the other branches of ca have thus made the identification of such practices a necessary analytic focus. 6. traditional ca and institutional ca ca has, since its beginning, noted that interaction within institutional settings differs in significant ways from everyday conversation. sacks frequently made reference to this fact in his lectures, while papers such as sacks, schegloff and jefferson (1974) and schegloff, jefferson and sacks (1977) note the existence of different systems of both turn-taking and repair within institutional forms of talk. it was not until atkinson and drew’s (1979) order in court that this variation was explored in detail rather than being simply mentioned in passing, however, with drew and heritage's (1992) talk at work following as one of the first books to entirely showcase key studies of institutional interaction from a conversation analytic perspective. heritage (1997) notes that the difference between institutional ca and traditional ca goes beyond a difference in the setting of the interactions being investigated, however: "there are, therefore, at least two kinds of conversation analytic research going on today, and though they overlap in various ways, they are distinct in focus. the first examines the social institution of interaction as an entity in its own right; the second studies the management of social institutions in interaction" (p. 223). in heritage's view, then, traditional ca is primarily concerned with goffman's (1983) concept of the "institutional order of interaction," and the way that talk-in-interaction both reflects and constitutes this order. institutional ca, on the other hand, finds itself more concerned with how particular institutions – medical, educational, legal, or otherwise – exist as relevant entities that shape and inform social interaction, and are both constructed and renewed by the talk itself. for example, ten have's (1991) institutional ca study of doctor-patient interaction was concerned far less with the individual and context-free practices for turn-taking employed by the participants, instead focusing on how the "asymmetry" in turn-taking that so frequently occurred between doctor and patient could be shown to be constituted in the interaction itself (by way of its institutional tenor) rather than as some pre-existing "social fact." 7. the invocation of institutionality 10 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 11 both traditional and institutional ca rely on the concept of procedural relevance as a means for determining the context of an interaction. the question of how to show the institutional character of an interaction to be demonstrably relevant to it has thus been a decades-long endeavor among analysts. drew and heritage (1992) note that given the "constraints" of a framework that treats context as locally and intra-interactionally produced, "analysts who wish to depict the distinctively 'institutional' character of some stretch of talk cannot be satisfied with showing that institutional talk exhibits aggregates and/or distributions of actions that are distinctive from ordinary conversation. they must rather demonstrate that the participants constructed their conduct over its course – turn by responsive turn – so as progressively to constitute … the occasion of their talk, together with their own social roles in it, as having some distinctively institutional character" (21). the avoidance of some a priori, assumed "institutionality" of an interaction is thus a priority for analysts working within institutional ca, who maynard and clayman (1991) claim carry this commitment so far as to be "concerned that using terms such as 'doctor's office', 'courtroom', 'police department', 'school room', and the like, to characterize settings ... can obscure much of what occurs within those settings" (p. 406-407). maynard and clayman argue here that this is the reason why conversation analysts interested in institutional interaction avoid relying on ethnographic knowledge about an institutional setting, and rely instead on participant orientations to these aspects of the context. (however, the avoidance of ethnographic data is a constraint that much work in institutional ca has adhered to less and less over the years). 8 yet the question remains as to what practices are used by participants to demonstrably orient to the institutional context of an interaction; and as schegloff (1992b) asks, "how does the fact that the talk is being conducted in some setting (e.g. 'the hospital') issue in any consequence for the shape, form, trajectory, content, or character of the interaction that the parties conduct? and what is the mechanism by which the context-so-understood has determinate consequences for the talk?" (111). drew and heritage (1992) answer this question in noting three ways that talk becomes "institutional" in the course of an interaction, each of which are procedurally relevant to participants. the first of these entails an orientation by the participants (or at least one participant) to some goal, task, or identity related to the particular institution being relevantly invoked. this "goal orientation" is, as heritage (1997) notes, a means for displaying not only the relevant institutionality of the interaction but also the "institutionally relevant identities" of the participants, e.g. doctor and patient (p. 163). the second means for introducing the institutionality of an interaction involves special constraints on the types of 8 this is especially true of maynard's work (e.g. 1984, 2003), despite the prior quote hinting at the contrary 11 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 12 contributions that are allowable to each participant over the course of the interaction, which may change from institutional task to institutional task. for example, in formal institutional settings (atkinson 1982) such as the courtroom or classroom, "there are specific reductions in the range of options and opportunities for action that are characteristic in conversation and they often involve specializations and respecifications of the interactional activities that remain (drew and heritage 1992: p. 26). the third marker of institutionality is the establishment of an institutionally-specific interpretive framework for the interaction, what drew and heritage refer to as "inferential frameworks and procedures" (p. 22). for example, the embodiment of a "professional" identity such as "doctor" or "judge" also entails the avoidance of expressing surprise, sympathy, agreement, or affiliation (the typically preferred responses in mundane conversational interaction) in response to the talk of "lay participants." the interpretation of these typically-disaligning actions as appropriate, unmarked responses for a participant embodying an institutional identity works to make relevant the institutional character of both the identity and the interaction itself. heritage (1997) notes that these three characteristics of institutional talk are most fruitfully probed for by the analyst in six arenas of interaction: "turn-taking organization, overall structural organization of the interaction, sequence organization, turn design, lexical choices, epistemological and other forms of symmetry" (p. 164). as these three criteria for institutionality show, then, the institutional context of an interaction is not something that exists prior to it, but is rather what heritage (1984) refers to as "ultimately and accountably talked into being" (p. 290). this idea of talking context into being is therefore much the same in institutional ca as it is in traditional ca, with heritage (1997) contrasting the approach to context in institutional ca with the aforementioned "bucket approach." as he notes, "it is fundamentally through interaction that context is built, invoked and managed, and that it is through interaction that institutional imperatives originating from outside the interaction are evidenced and made real and enforceable for the participants" (p. 163). heritage here gives an institution-specific example of this approach to context by way of an emergency call to the police, and describes the management of context as visible in the ways that participants "are managing their interaction as an ‘emergency call’ on a ‘policeable matter’. we want to see how the participants co-construct it as an emergency call, incrementally advance it turn by turn as an emergency call, and finally bring it off as having been an emergency call" (p. 163). as the discussion above shows, there is considerable variation between speaker practices across institutional and everyday forms of talk-in-interaction (for example, the constraints on what types of contributions are allowable to each participant within an interaction, or the institution-specific frames for interpretation utilized by the interactants). this variation is what heritage and greatbatch (1991) refer to as "contributing to a unique 'fingerprint' for each institutional form of interaction – the 'fingerprint' being comprised of a set of 12 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 13 interactional practices differentiating each form both from other institutional forms and from the baseline of mundane conversational interaction itself" (p. 9596). it is thus participant orientations to the practices that comprise these "fingerprints" that make relevant an institutional context for the participants. however, an understanding of what comprises these fingerprints to begin with is something that may entail the acquisition of specialized forms of knowledge particular to an institutional setting. likewise, the ability of the analyst to recognize these practices as talking institutions into being would also require such expanded arenas of knowledge. there are numerous ways that participants talk institutional contexts into being that require access to these types of knowledge, as well as to particular "common-sense" understandings of the lifeworld of the participants. for example, heritage and sefi (1992) highlight the relevance of the next-turn proof procedure in showing how an utterance that may appear to be a casual observation – "he’s enjoying that isn’t he" – may elicit responses that reflect both the institutional tenor of the interaction and the division of labor among the responding participants. within this particular example, the observation is delivered by a health visitor to the parents of the baby that she is evaluating. the father of the child orients to the observational character of the health visitor's turn at talk and provides an aligning response ("yes, he certainly is"). the mother’s response, however, shows a notably different orientation ("he’s not hungry cus he’s just had ‘iz bottle"), and displays what the authors describe as a notable "defensiveness" as it rejects an unstated inference of the health visitor's remark: that the baby is chewing on something because he is hungry. the mother's turn at talk thus displays an orientation to one of the institutionally-ascribed goals of the health worker, to evaluate whether the needs of the child (e.g. being adequately fed) are being met. given the notably distinct orientations to the health worker's talk by the mother and father, the analysts conclude that yet another relevant aspect of the interaction (for analyst and participant alike) is the division of labor within the family, "in which the mother is treated as having the primary responsibility for her baby (reflected in her defensiveness), while the father, with less responsibility, can take a more relaxed and ‘innocent’ view of things" (heritage 1997: p. 232). notably, both the institutional task of the health worker and an awareness of the expectations of mothers within the familial division of labor – each of which contribute to the context of the interaction – make use of ethnographicallyoriented forms of knowledge and insights that are acknowledged far more frequently within institutional ca than traditional ca. this is not to say that traditional ca doesn't make similar use of the category predicates associated with other types of identities: orientations to familial roles can show up in dinner 13 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 14 conversations, for example, where "mom" might have cooked the meal and may respond "defensively" to a criticism of the food. 9 in such a case, both the institutional task of the mother/cook and an awareness of the expectations of mothers within the familial division of labor can potentially contribute to the context of the interaction, just as in the example of the health worker described above. the use of ethnographically-oriented forms of knowledge and insights may thus be crucial to understanding the context-sensitive practices in operation even in traditional ca, particularly in single-case analyses where descriptions of context-free practices may be heavily reliant upon the analyst's understanding of context-sensitive practices. as mentioned earlier, however, traditional ca's general avoidance of framing either data or findings as sensitive to ethnographic issues leads to a perceivable difference between the analytic practices across traditional and institutional ca. another fruitful site for how the forms of knowledge discussed above may invoke an institutional context is a speaker's lexical choices. institution-specific uses of lexical items within institutional contexts were noticed early on by sacks, who notes the relevance of how the term "cop" is used to describe police officers in everyday conversation while the term ‘police officer’ is used while giving evidence in court (1979), or how members of organizations refer to themselves as "we" rather than "i" (1992). heritage (1997) echoes these findings by noting that "a clear way in which speakers orient to institutional tasks and contexts is through their selection of descriptive terms" (p. 173-174). to illustrate this point, he shows how the self-identification of a school employee in the opening sequence of a telephone call with a student's mother (by using a "last name + organizational identification" format rather than a more mundane "first + last name") allows the mother to "identify the phone call as a 'business call' and, specifically, a 'call about school business'" (p. 175). schegloff (1987) and (1992b) also draw upon the import of lexical choice in discussing how the "pointed use of a technical or vernacular idiom" (such as the use of "hematoma" rather than the lay term "bruise") makes relevant a form of knowledge or state of expertise specific to an institutional identity. such practices thus potentially display the relevance of that identity to the participants within a particular interaction. he here draws on cicourel (1987), who argues that technical medical terms "anchor within the interaction the relevance for the participants of the medical cast of the setting and of the participants...and invokes it within the interaction" (schegloff 1992b, p. 197). while schegloff notes that the analyst working in institutional ca would not simply claim the relevance of the institutional tenor of such an interaction based on "extrinsic ethnographic 9 of course, if we treat the family as an institution (a common understanding across the social sciences) then we see how "institutional identities" may become just as relevant to interactions at the dinner table as they are to interactions in the courtroom. 14 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 15 grounds," he does note in a footnote within the 1992b paper that "ethnographic research may, of course, [be] necessary to enable the analyst to recognize the sense and import of such terms as display[ing] the relevance of some aspect of context, or to recognize that seemingly ordinary words have such an import, but the relevance of whatever has been learned through fieldwork (or in any other manner) must be warranted as relevant to the participants by reference to details of the conduct of the interaction" (223). that ca may need to be "informed" by the analyst's ethnographic practices, then, is something discussed far more openly within institutional ca than traditional ca. this lack of acknowledgement within traditional ca has contributed to the common understanding of traditional ca as being non-reliant on ethnographic (or "member's") forms of knowledge, as seen in arminen's (2005) claim that "ca studies do not generally rely on ethnographic knowledge, but the analysis of some institutional settings may require contextual knowledge in order to make sense of realms distinct from everyday life" (p. 1). maynard (2003) makes a contrasting point to this observation, however, arguing instead that traditional ca draws upon this same range of ethnographic information; the difference between traditional ca and institutional ca is only that the latter openly admits to the practice. as he argues, "ethnographic knowledge – an insider's understanding of terms, phrases, and courses of action – is something that ca regularly draws upon when displaying and analyzing a particular excerpt" (p. 74). these are points that throw into sharp relief one of the perceived distinctions between traditional ca and institutional ca: the acknowledgment that distinct forms of knowledge are employed by both participant and analyst in navigating and formulating the context of an interaction. arminen (2005), for example, notes that the traditional understanding of context as a demonstrably oriented-to feature of interaction "trades on the analyst’s taken-for-granted competence in presupposing an argument that formulates the context-relevant features of interaction to the object of scrutiny" (35). arminen claims that within sequential ca, there is the assumption that "an analyst is automatically competent to identify the context-relevant features of an interaction. … if one cannot point to the relevant presence of the ... context in the interaction, the problem may either be that the context is irrelevant for the accomplishment of that action or that the analyst’s argument has been inadequate and has not allowed us to notice the relevance of context" (p. 35-36, emphasis in original). arminen goes so far as to suggest that "if we are to study institutional interaction in its own right we have to revise schegloff's methodological policy" (p. 37). he supports this claim by drawing on an example of one of schegloff's (1991) critiques of zimmerman's (1984) study of emergency calls as overestimating the institutional relevance of the practices studied therein (the calltaker's use of an "interrogative series" of insertion sequences), thereby missing "the potentially general relevance of insertions to sequences of this type" (1991: 59). however, as arminen notes, schegloff's critique is couched in what heritage 15 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 16 (1997) refers to as the different goals of traditional and institutional ca: "the social institution of interaction as an entity in its own right," and "the management of social institutions in interaction," respectively. as arminen states, "when the task of the analysis is not only to describe sequential patterns of interaction, but to identify and explicate the ways in which interactional activities contribute to the accomplishment of institutional tasks, then the analyst's ability to connect the interactional patterns to the institutional activities becomes essential and makes relevant the analyst's context-sensitive understanding of the institutional tasks" (2005, 37). whereas the earlier sacks, schegloff, and jefferson (1974) placed a priority on the context-free practices of talk (in their previously mentioned claim that "it is undesirable to have to know or characterize such situations for particular conversations in order to investigate them"), arminen here suggests that attention to those context-sensitive aspects of an interaction may be far more heavily weighted within institutional ca. however, if we return to the earlier example of a dinner conversation in which the family-institutional role "mom" is made relevant to the ongoing tasks and actions of the interaction, we can see how context-sensitive aspects of an interaction may be similarly relevant to work within traditional ca. 10 in addition to the traditional-institutional differences mentioned above, arminen also argues that particular institutional fields and professions may have their own sets of beliefs and theories of social interaction (what peräkylä and vehviläinen 2003 refer to as "interaction ideologies" or "stocks of professional knowledge,") and that these too form potential aspects of context that analysts need to be aware of to paint a full picture of the context of an interaction. arminen thus leaves open the potential that similar types of ideologies exist in contextually relevant ways to participants in non-institutional settings, an argument that has been frequently seen elsewhere within sociocultural linguistics, though has yet to be a relevant aspect of traditional ca. 8. concluding remarks though recent scholarship in ca has recognized the existence of such named "sub-varieties" as institutional ca and feminist ca, relatively little work has explored the variation in analytic practice and scope seen across these branches in any detail. as the discussion above illustrates, the notion of "context" within ca provides one fruitful area for comparison. though focused here on comparing only traditional ca and institutional ca, work within each of the branches 10 from this, it might be suggested that "everyday life" is treatable as a kind of institution, or perhaps even a variety of them. 16 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 17 outside of traditional ca – in engaging with a broader understanding of context that still seeks to align with the participants' own, endogenous understandings of their everyday interactions – problematizes the lack of concern within traditional ca to consider the analyst's own ethnographic or member's epistemologies as potentially relevant to the analysis. one direction for the future of the field may thus lie in bridging this gap between the analytic approaches adopted across the different branches of ca. as this paper also illustrates, there is (and should not be) no single means for doing ca or for approaching the concept of "context" within it, despite the occasional use of such intimidating modifiers as "schegloffian" to refer to the "traditional" branch of the framework. rather, ca may be best understood as a methodological approach with core tenets, such as the understanding that context is produced locally and intersubjectively as an interactional resource that is procedurally relevant to the surrounding talk. 17 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 18 references arminen, ilkka. 2005. institutional interaction: studies of talk at work. aldershot: ashgate. atkinson, j. maxwell and paul drew. 1979. order in court: the organisation of verbal interaction in judicial settings. london: macmillan. bar-hillel, yehoshua. 1954. "indexical expressions." mind 63: 359–379. bauman, richard and joel sherzer. 1974. explorations in the ethnography of speaking. cambridge: cambridge university press. billig, michael. 1999. "whose terms? whose ordinariness? rhetoric and ideology in conversation analysis." discourse & society 10: 543-58. billig, michael. 1999. "conversation analysis and the claims of naivety." discourse & society 10: 572-76. blommaert, jan. 2001. "context is/as critique." critique of anthropology 21(1): 13–32 boden, d. and zimmerman, d.h. (eds.) .1991. talk and social structure. cambridge: polity press. briggs, charles. 1997. "notes on a 'confession': on the construction of gender, sexuality, and violence in an infanticide case." pragmatics 7(4): 519–46. bucholtz, mary. 2003. "theories of discourse as theories of gender: discourse analysis in language and gender studies." in janet holmes and miriam meyerhoff (eds.) the handbook of language and gender. oxford, blackwell: 43-68. carnap, rudolf. 1942. introduction to semantics. cambridge, mass: harvard university press. cicourel, aaron v. 1981. "notes on the integration of microand macro-levels of analysis." in karin knorr-cetina and aaron v. cicourel (eds.) advances in social theory and methodology: toward an integration of microand macrosociologies. routledge, boston: 51-80. cicourel, aaron v. 1987. "the interpenetration of communicative contexts: examples from medical encounters." social psychology quarterly 50: 217226. (revised version in a. duranti and c. goodwin (eds.) (1992) rethinking context: language as an interactive phenomenon, cambridge university press: 291-310. drew, paul and john heritage. 1992. "analyzing talk at work: an introduction." in paul drew and john heritage (eds.), talk at work, cambridge, cambridge university press: 3-65. edley, nigel. 2001. "conversation analysis, discursive psychology and the study of ideology: a response to susan speer." feminism and psychology 11(1): 136–40. 18 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 19 garfinkel, harold, harvey sacks. 1970. "on formal structures of practical action." in john c mckinney & e.a. tiryakian (eds.) theoretical sociology: perspectives and developments. new york, appleton-century-crofts: 338-66 gazdar, gerald. 1979. pragmatics: implicature, presupposition, and logical form. london: academic press. goffman, erving. 1959. the presentation of self in everyday life. new york, doubleday, anchor books. goffman, erving. 1983. "the interaction order." american sociological review 48: 1-17. goodwin, marjorie harness. 1990. he-said-she-said: talk as social organization among black children. bloomington, in: indiana university press. gumperz, john j. and dell hymes. 1972. directions in sociolinguistics: the ethnography of communication. new york: holt, rinehard and winston. hak, tony. 1995. "ethnomethodology and the institutional context." human studies 18: 109–37. hammersley, martin. 2003. "conversation analysis and discourse analysis: methods or paradigms?" discourse & society 14(6): 751-781. have, paul ten. 1990. "methodological issues in conversation analysis." bulletin de méthodologie sociologique 27 (june): 23-51 have, paul ten. 1991. "talk and institution: a reconsideration of the 'asymmetry' of doctor patient interaction." in deirdre boden & don h. zimmerman (eds.) talk and social structure: studies in ethnomethodology and conversation analysis. cambridge, polity press: 138-63. retrieved from http://www.paultenhave.nl/mica.htm heritage, john. 1984. garfinkel and ethnomethodology. cambridge: polity press heritage, john. 1997. "conversation analysis and institutional talk: analysing data." in david silverman (ed.) qualitative research: theory, method and practice. london, sage: 161-82. heritage, john. 2008. "conversation analysis as social theory." in bryan turner (ed.) the new blackwell companion to social theory. oxford, blackwell: 300-320. heritage, john and david greatbatch. 1991. "on the institutional character of institutional talk: the case of news interviews." in deirdre boden, don h. zimmerman (eds.) talk and social structure: studies in ethnomethodology and conversation analysis. cambridge, polity press: 93-137 heritage, john, sue sefi. 1992. "dilemmas of advice: aspects of the delivery and reception of advice in interactions between health visitors and first time mothers." in paul drew and john heritage (eds.) talk at work. cambridge, cambridge university press: 359-419. hutchby, ian and robin wooffitt. 1998. conversation analysis: principles, practices and applications. blackwell publishers inc. hymes, dell. 1972. "models of the interaction of language and social life." in gumperz and hymes (eds.) directions in sociolinguistics: the ethnography of communication. new york, holt, rinehart and winston: 35-71 19 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 20 hymes, dell. 1974. foundations in sociolinguistics: an ethnographic approach. philadelphia: university of pennsylvania press. jakobson, roman. 1960. "concluding statements: linguistics and poetics." in style in language (ed.) thomas a. sebeok. cambridge, mit press: 350-377. kitzinger, celia. 2000. "doing feminist conversation analysis." feminism and psychology 10(2): 163–93. labov, william. 1966. the social stratification of english in new york city. washington, d.c., center for applied linguistics. labov, william. 1972a. language in the inner city: studies in the black english vernacular. philadelphia, university of pennsylvania press. labov, william. 1972b. sociolinguistic patterns. philadelphia, university of pennsylvania press. leech, geoffrey n. 1983. principles of pragmatics. london: longman. levinson, stephen. 1983. pragmatics. cambridge: cambridge university press. lerner, gene h., celia kitzinger. 2007. "introduction: person-reference in conversation analytic research." discourse studies 9: 427-432. lynch, michael. 1985. art and artifact in laboratory science: a study of shop work and shop talk. london: routledge & kegan paul malinowski, bronislaw. 1923. "the problem of meaning in primitive languages, in the meaning of meaning." in c.k. ogden and i.a. richards (eds.) new york, harcourt, brace and world, inc.: 296-336 malinowski, bronislaw. 1935. coral gardens and their magic, 2 vols. london: allen and unwin. maynard, douglas w. 2003. bad news, good news: conversational order in everyday talk and clinical settings. chicago, university of chicago press maynard, douglas w. and steven e. clayman. 1991. "the diversity of ethnomethodology." annual review of sociology 17: 385-418. moerman, michael. 1988. talking culture: ethnography and conversational analysis. philadelphia: university of pennsylvania press morris, c.w. 1938. "foundations of the theory of signs." in o. neuerath, r. carnap, and c. morris (eds.) international encyclopedia of unified science. chicago: university of chicago press: 77-138. ochs, elinor. 1979. "what child language can contribute to pragmatics." in e. ochs & b. schieffelin (eds.) developmental pragmatics. new york: academic press: 1-17. ochs, elinor, emanuel schegloff, and sandy thompson. 1996. interaction and grammar. cambridge: cambridge university press. peräkylä anssi, sanna vehviläinen. 2003. "conversation analysis and the professional stocks of interactional knowledge." discourse & society 14(6): 727-50 psathas, george. 1995. "'talk and social structure' and 'studies at work'." human studies 18: 139–55. 20 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 approaches to "context" within conversation analysis 21 raclaw, joshua. 2010. "member knowledge and ethnographic insight: the relevance of analyst knowledge in doing conversation analysis." clic conference, ucla: may 8, 2010. sacks, harvey. 1979. "hotrodder: a revolutionary category." in george psathas (ed.) everyday language: studies in ethnomethodology. new york, irvington: 7-14 sacks, harvey. 1992. lectures on conversation. 2 vols. edited by gail jefferson with introductions by emanuel a. schegloff. oxford, basil blackwell. sacks, harvey, emanuel a. schegloff, and gail jefferson. 1974. "a simplest systematics for the organization of turn taking for conversation." language 50: 696-735. schegloff, emanuel a. 1972. "notes on a conversational practice: formulating place." in david sudnow (ed.). studies in social interaction. new york, free press: 75-119. schegloff, emanuel a. 1987. "between macro and micro: contexts and other connections." in j. alexander, et al. (eds.) the micro-macro link. berkeley and los angeles, university of california press: 207-34. schegloff, emanuel a. 1988. "discourse as an interactional achievement ii: an exercise in conversation analysis." in: d. tannen (ed.) linguistics in context: connecting observation and understanding. norwood, n.j., ablex. schegloff, emanuel a. 1991. "reflections on talk and social structure." in: boden, deirdre, don h. zimmerman, eds. talk and social structure: studies in ethnomethodology and conversation analysis. cambridge: polity press: 4470. schegloff, emanuel a. 1992a. "introduction." in h. sacks, lectures on conversation. 2 vols. edited by gail jefferson with introductions by emanuel a. schegloff. oxford: basil blackwell. schegloff, emanuel a. 1992b. "in another context." in a. duranti and c. goodwin (eds.) rethinking context: language as an interactive phenomenon. cambridge, cambridge university press: 193-227 schegloff, emanuel a. 1997. "whose text? whose context?" discourse & society 8: 165-87. schegloff, emanuel a. 1998. "reply to wetherell." discourse & society 9: 41316. schegloff, emanuel a. 1999. "'schegloff’s texts' as 'billig’s data': a critical reply." discourse & society 10: 558-72 schegloff, emanuel a. 1999. "naivety vs. sophistication or discipline vs. selfindulgence: a rejoinder to billig." discourse & society 10: 577-82. schegloff, emanuel a. 2002. "conversation analysis, then and now." plenary address for the inaugural session of the section-in-formation on ethnomethodology and conversation analysis of the american sociological association, chicago, august 19, 2002. schegloff, emanuel a. 2007. "a tutorial on membership categorization." journal of pragmatics 39: 462-82 21 raclaw: approaches to "context" within conversation analysis published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 22 schegloff, emanuel a. 2009. "one perspective on conversation analysis: comparative perspectives." in jack sidnell (ed.) conversation analysis: comparative perspectives. cambridge: cambridge university press: 357-406 schegloff, emanuel a. and harvey sacks. 1973. "opening up closings." semiotica 8: 289-327. schegloff, emanuel a., gail jefferson, harvey sacks. 1977. "the preference for self-correction in the organization of repair in conversation." language 53: 361-82. searle, john. 1986. "introductory essay; notes on conversation." in d. ellis and w. donahue (eds.) contemporary issues in language and discourse processes. hillsdale, nj, lawrence erlbaum associates: 7-19. silverman, david. 1998. harvey sacks: social science and conversation analysis. oxford: oxford university press. speer, susan. 1999. "feminism and conversation analysis: an oxymoron?" feminism and psychology 9(4): 471–478. speer, susan. 2001a. "reconsidering the concept of hegemonic masculinity: discursive psychology, conversation analysis and participants’ orientations." feminism and psychology 11(1): 107–35. speer, susan. 2001b. "participants’ orientations, ideology and the ontological status of hegemonic masculinity: a rejoinder to nigel edley." feminism and psychology 11(1): 141–4. stokoe, elizabeth h. and janet smithson. 2001. "making gender relevant: conversation analysis and gender categories in interaction." discourse & society 12: 217–44. watson, rod. 1997. "some general reflections on ‘categorization’ and ‘sequence’ in the analysis of conversation." in hester, s., peter eglin (eds.) culture in action: studies in membership categorization analysis. washington, d.c., university press of america: 49-76. wetherell, m. 1998. "positioning and interpretative repertoires: conversation analysis and poststructuralism in dialogue." discourse & society 9: 387-412 zimmerman, don h. 1984 "talk and its occasion: the case of calling the police." in d. shiffren (ed.) meaning, form and use in context: linguistic applications. washington, d.c., georgetown university press: 210-28 22 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/1 doi: https://doi.org/10.25810/qbrs-0970 colorado research in linguistics 6-2009 approaches to "context" within conversation analysis joshua raclaw recommended citation cril style sheet “code switching” in sociocultural linguistics colorado research in linguistics. june 2006. vol. 19. boulder: university of colorado. © 2006 by chad nilep. “code switching” in sociocultural linguistics chad nilep university of colorado, boulder this paper reviews a brief portion of the literature on code switching in sociology, linguistic anthropology, and sociolinguistics, and suggests a definition of the term for sociocultural analysis. code switching is defined as the practice of selecting or altering linguistic elements so as to contextualize talk in interaction. this contextualization may relate to local discourse practices, such as turn selection or various forms of bracketing, or it may make relevant information beyond the current exchange, including knowledge of society and diverse identities. introduction the term code switching (or, as it is sometimes written, code-switching or codeswitching)1 is broadly discussed and used in linguistics and a variety of related fields. a search of the linguistics and language behavior abstracts database in 2005 shows more than 1,800 articles on the subject published in virtually every branch of linguistics. however, despite this ubiquity – or perhaps in part because of it – scholars do not seem to share a definition of the term. this is perhaps inevitable, given the different concerns of formal linguists, psycholinguists, sociolinguists, philosophers, anthropologists, etc. this paper will attempt to survey the use of the term code switching in sociocultural linguistics and suggest useful definitions for sociocultural work. since code switching is studied from so many perspectives, this paper will necessarily seem to omit important elements of the literature. much of the work labeled “code switching” is interested in syntactic or morphosyntactic constraints on language alternation (e.g. poplack 1980; sankoff and poplack 1981; joshi 1985; di sciullo and williams 1987; belazi et al. 1994; halmari 1997 inter alia). alternately, studies of language acquisition, second language acquisition, and language learning use the term code switching to describe either bilingual speakers’ or language learners’ cognitive linguistic abilities, or to describe classroom or learner practices involving the use of more than one language (e.g. romaine 1989; cenoz and genesee 2001; fotos 2001, inter alia). these and other studies seem to use code as a synonym for language variety. alvarez-cáccamo 1 my personal preference is to spell code switching as two words, with white space between them, a practice i will generally follow throughout this paper. original spelling will be preserved in quotations and when paraphrasing scholars who routinely use an alternate form. 1 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 2 (2000) argues that this equation may obscure certain interactional functions of such alternation. practically all work on “code-switching,” or changing codes, has been based on a strict identification between the notions of “code” and “linguistic variety,” be that a language, dialect, style, or prosodic register. however, this structural focus fails to convincingly explain certain conversational phenomena relative to the relevance or significance (or lack of relevance) of alternations between contrasting varieties. [alvarez-cáccamo 2000:112; my translation] certainly, the study of language alternation has been fruitful over the past several decades. the identification of various constraints, though sometimes controversial, has inspired a great deal of work in syntax, morphology, and phonology. a structural focus has been similarly constructive for production models (e.g. azuma 1991) or as evidence for grammatical theory (e.g. macswann 2000; jake, myers-scotton and gross 2002). by ignoring questions of function or meaning, though, this structural focus fails to answer basic questions of why switching occurs.2 auer (1984) warns, “grammatical restrictions on codeswitching are but necessary conditions” (2); they are not sufficient to describe the reason for or effect of a particular switch. if linguists regard code switching simply as a product of a grammatical system, and not as a practice of individual speakers, they may produce esoteric analyses that have little importance outside the study of linguistics per se, what sapir called “a tradition that threatens to become scholastic when not vitalized by interests which lie beyond the formal interest in language itself” (1929:213). this paper is thus positioned within the discipline of sociocultural linguistics, an emerging (or one might say, revitalized) approach to linguistics that looks beyond formal interests, to the social and cultural functions and meanings of language use. periodically over the last century, linguists have proposed to bring their own studies closer to other fields of social inquiry. in 1929, edward sapir urged linguists to move beyond diachronic and formal analyses for their own sake and to “become aware of what their science may mean for the interpretation of human conduct in general” (1929:207). he suggested that anthropology, sociology, psychology, philosophy and social science generally would be enriched by drawing on the methodologies as well as the findings of linguistic research. he also exhorted linguists to consider language within its broader social setting. 2 woolard (2004) suggests that the basic question should be not why speakers make use of the various forms available to them, but why speakers would not make use of all available forms. thus she suggests, “it could be argued that linguists, with their focus on constraints against rather than motivations for codeswitching, do ask this alternative question” (91). 2 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 3 it is peculiarly important that linguists, who are often accused, and accused justly, of failure to look beyond the pretty patterns of their subject matter, should become aware of what their science may mean for the interpretation of human conduct in general. whether they like it or not, they must become increasingly concerned with the many anthropological, sociological, and psychological problems which invade the field of language. [sapir 1929:214] sapir was not alone in his hopes for a more socially engaged linguistics. indeed the development of sociolinguistics and psycholinguistics during the 1930s-1950s suggests that, at least for some linguists, social interaction and human cognition were as important as the forms and structures of language itself. nonetheless, by the 1960s some scholars once again felt the need to argue for a more socially engaged linguistics. in a special issue of american anthropologist, hymes (1964) lamented that the socially integrated linguistics sapir had called for was disappearing. hymes and others worried that new formal approaches, as well as the push for linguistics as an autonomous field, threatened to once again isolate linguists. at the same time, though, the growth of ethnolinguistics and sociolinguistics offered a venue for the socially engaged linguistics sapir had called for four decades earlier. four more decades have passed, and once again scholars are calling for a revitalization of socially and culturally oriented linguistic analysis. bucholtz and hall (2005) position their own work on language and identity as what they call sociocultural linguistics, “the broad interdisciplinary field concerned with the intersection of language, culture, and society” (5). just as hymes (1964) worried that linguistics had been bleached of its association with the study of human interaction in the wake of formalist studies, bucholtz and hall point out that sociolinguistics has in turn been narrowed to denote only specific types of study. sociocultural linguistics is thus suggested as a broader term, to include sociolinguistics, linguistic anthropology, discourse analysis, and sociology of language, as well as certain streams of social psychology, folklore studies, media studies, literary theory, and the philosophy of language. what follows is a brief survey of work on the topic of code switching within sociocultural linguistics, followed by my own suggested definition for the term. i hope this definition will serve as a basis and context for sociocultural discussions of the contextualizing functions of language alternation and modulation. 1. foundational studies 1.1. early studies: the emergence of code switching the history of code switching research in sociocultural linguistics is often dated from blom and gumperz’s (1972) “social meaning in linguistic structures” (e.g. myers-scotton 1993; rampton 1995; benson 2001). this work is certainly 3 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 4 important and influential, not least for introducing the terms situational and metaphorical switching (see below). however, by 1972 the term “code switching” was well attested in the literature, and several studies in linguistic anthropology and sociolinguistics prefigured later code switching research in sociocultural linguistics. below, i survey some important early work. one of the earliest american studies in linguistic anthropology to deal with issues of language choice and code switching was george barker’s (1947) description of language use among mexican americans in tucson, arizona. in addition to his analysis of the economic relations, social networks, and social geography of tucson residents, barker sought to answer the question, “how does it happen, for example, that among bilinguals, the ancestral language will be used on one occasion and english on another, and that on certain occasions bilinguals will alternate, without apparent cause, from one language to another?” (1947:18586). barker suggested that interactions among family members or other intimates were most likely to be conducted in spanish, while formal talk with angloamericans was most likely to use the medium of english (even when all parties in the interaction were able to understand spanish). in less clearly defined situations, language choice was less fixed, and elements from each language could occur. further, barker proposed that younger people were more apt to use multiple languages in a single interaction than were their elders, and that the use of multiple varieties was constitutive of a local tucson identity. an important base for code switching research in the field of linguistics is uriel weinreich’s (1953) languages in contact. one of those inspired by weinreich’s book was hans vogt, whose “language contacts” (1954) is cited as the first article to use the term “code-switching” in the field of linguistics (alvarez-cáccamo 1998; benson 2001). weinreich was interested to describe the effect of language contact on languages, in addition to describing the activities of bilingual speech communities. he suggested that barker’s (1947) description of tucson was insufficient, since it listed only four speech situations: intimate, informal, formal, and inter-group discourse. weinreich argued that barker’s taxonomy was “insufficiently articulated” (87) to describe all potential organizations of bilingual speech events. he contended that anthropology should look to linguistics – particularly to structuralism – in order to properly describe the practice of bilingual speech, and the language acquisition/socialization process that takes place in bilingual communities. 4 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 5 weinreich’s description of switching codes3 suggested that bilingual individuals possess two separate linguistic varieties, which (ideally) they employ on separate occasions. he suggested that frequent alternation, such as that barker described among tucson youth, was a product of poor parenting. regular code switchers, weinreich speculated, “in early childhood, were addressed by the same familiar interlocutors indiscriminately in both languages” (74).4 this indiscriminate use differed from the ideal bilingual of weinreich’s imagination. vogt’s (1954) article, though very much inspired by weinreich (1953), is much less apprehensive about bilingual code switching. code-switching in itself is perhaps not a linguistic phenomenon, but rather a psychological one, and its causes are obviously extralinguistic. but bilingualism is of great interest to the linguist because it is the condition of what has been called interference between languages. [vogt 1954:368] vogt assumes that code switching is not only natural, but common. he suggests that all languages – if not all language users – experience language contact, and that contact phenomena, including language alternation, are an important element of language change. the phenomenon of diglossia, first described by ferguson (1959), and later refined by fishman (1967), is another precursor to linguistic analyses of code switching. ferguson defined diglossia as the existence of a “divergent, highly codified” (1959:336) variety of language, which is used only in particular situations. although ferguson limited diglossia to varieties of the same language, fishman (1967) described similar functional divisions between unrelated languages. neither ferguson nor fishman cite examples of alternation between varieties within a single interaction or discourse. however, their descriptions of diglossia bear on the notion of situational switching. furthermore, fishman, citing an unpublished paper by blom and gumperz, mentions that varieties may be employed for humor or emphasis in a process of metaphorical switching (fishman 1967:36). thus, fishman’s account of diglossia at least seems to have been 3 the notion of “switching codes” appears to have been borrowed from information theory. weinreich refers to fano 1950, a paper also referenced by jakobson (1971a [1953], 1971b [1961]; jakobson and halle 1956) in his discussions of code switching. fuller exploration of these links is unfortunately beyond the scope of the present paper. see alvarez-cáccamo (1998, 2000) for more detail. 4 for discussion of the one-person-one-language ideology in language acquisition see romaine (1989). 5 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 6 inspired by the nascent theory of situational and metaphorical switching (blom and gumperz 1972; see below).5 erving goffman (1979, 1981) described footing as a process in interaction similar to some functional descriptions of code switching. indeed, goffman cites several of gumperz’s descriptions of code switching as examples of footing. the difference he draws between his own theory of footing and gumperz’s and others’ descriptions of code switching is a formal one. whereas code switching (at least for goffman) necessarily includes a shift from one language to another,6 footing shifts may also be indicated in a variety of ways. even so, goffman writes, “for speakers, code switching is usually involved” in footing shifts, “and if not this then at least the sound markers that linguists study: pitch, volume, rhythm, stress, [or] tonal quality” (goffman 1981:128). for goffman, footing is the stance or positioning that an individual takes within an interaction. within a single interaction – even within a short span of talk – an individual can highlight any number of different roles. goffman suggests that changes in purpose, context, and participant role are common in interaction, and offers footing as a useful theory of the multiple positions taken by parties to talk in interaction. during the course of an interaction, an individual is likely to display a number of different stances; much of goffman’s discussion of footing is thus dedicated to switches in footing. alternating languages, among other linguistic markers, can serve to mark these shifts in context or role. 1.2. gumperz: code switching and contextualization perhaps no sociocultural linguist has been more influential in the study of code switching than john j. gumperz. his work on code switching and contextualization has been influential in the fields of sociolinguistics, linguistic anthropology, and the sociology of language. much of gumperz’s early work was carried out in northern india (gumperz 1958, 1961, 1964a, 1964b), focused on hindi and its range of dialects. gumperz 1958 describes three levels – village dialects, regional dialects, and standard hindi – each of which may be comprised of numerous varieties, and which serve different functions. gumperz writes, “most male residents, especially those who travel considerably, speak both the village and the regional dialect. the former is used at home and with other local 5 fishman also credits gumperz for expanding the notion of diglossia to include multilingual societies. however, studies fishman cites as diglossia were labeled by gumperz as code switching. 6 it is far from clear that early code switching research assumed such strict separation of languages. blom and gumperz 1972, for example, focus on two dialects of spoken norwegian. similarly, fishman states explicitly, “a theory [of diglossia] which tends to minimize the distinction between languages and varieties is desirable for several reasons” (1967:33). 6 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 7 residents; the latter is employed with people from the outside” (1958:669). thus the relationship between speakers affects the choice of language variety. the idea that linguistic form is affected by setting and participants as well as topic was influenced in part by ervin-tripp (1964). her definitions of setting, topic, and function provide an important base for the work of gumperz and others. her study of bilingual japanese-born women living in the united states observed considerable correlation between language choice and discourse content, providing an example of “semantic” analysis of language choice that, while influential (e.g. myers-scotton 1993), would be criticized as only partial and approximate (e.g. auer 1984, 1995). in 1963, while working with the institute of sociology at oslo university, gumperz met jan-petter blom (dil 1971). together, blom and gumperz undertook a study of verbal behavior in hemnesberget, a small settlement of about 1,300 people in northern norway. gumperz (1964b) compared the use of two dialects, standard literary bokmål and local ranamål, in hemnesberget to the use of standard and local dialects of hindi in northern india. in each population, the local dialect appeared more frequently in interaction with neighbors, while the standard dialect was reserved for communication across “ritual barriers” (148) – barriers of caste, class, and village groupings in india, and of academic, administrative, or religious setting in norway. on the basis of these comparisons, gumperz argued that verbal repertoire is definable in social as well as linguistic terms. distinct repertoires are identified in terms of participants, setting, and topic, and then described in terms of phonological and morphological characteristics. blom and gumperz (1972) expanded the analysis of the functions of bokmål and ranamål in hemnesberget in what has come to be a touchstone in code switching research. they described bokmål and ranamål as distinct codes, though not distinct languages. the codes are distinguished by extensive though slight phonological, morphological and lexical differences, as well as native speakers’ belief that the two varieties are separate, and tendency to maintain that separation of form. blom and gumperz asked why, despite their substantial similarities, and the fact that most speakers commanded both varieties, bokmål and ranamål were largely maintained as separate. “the most reasonable assumption,” they argued, “is that the linguistic separateness between dialect and standard… is conditioned by social factors” (417). thus, each variety was seen as having low level differences in form, as well as somewhat distinct social functions. blom and gumperz posited that social events, defined in terms of participants, setting, and topic, “restrict the selection of linguistic variables” (421) in a manner that is somewhat analogous to syntactic or semantic restrictions. that is, in particular social situations, some linguistic forms may be more appropriate than others. among groups of men greeting each other in workshops along the fjord, the variety of language used differed from that used by teachers presenting text material in the public school, for example. it is important to recognize that 7 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 8 different social events may, for example, involve the same participants in the same setting when the topic shifts. thus, teachers reported that they treated lecture versus discussion within a class as different events. while lectures were (according to teachers’ reports) delivered in the standard bokmål, a shift to the regional ranamål was used to encourage open debate. blom and gumperz call this type of shift, wherein a change in linguistic form represents a changed social setting, situational switching (424). the definition of metaphorical switching relies on the use of two language varieties within a single social setting. blom and gumperz describe interactions between clerks and residents in the community administration office wherein greetings take place in the local dialect, but business is transacted in the standard. in neither of these cases is there any significant change in definition of participants’ mutual rights and obligations. … the choice of either (r) or (b)… generates meanings which are quite similar to those conveyed by the alternation between ty and vy in the examples from russian literature cited by friedrich [1972]. we will use the term metaphorical switching for this phenomenon. [blom and gumperz 1972:425] blom and gumperz suggest that the use of local (r) phrases in a standard (b) conversation allude to other social events in which the participants may have been involved. this allusion lends some connotative meaning, such as confidentiality, to the current event, without changing the topic or goal. the notions of situational and metaphorical switching were taken up by a great many sociolinguists, linguistic anthropologists, etc. whereas blom and gumperz identified ranamål and bokmål as “codes in a repertoire” (414) and went to some pains to describe the formal differences between the two, many subsequent scholars have been content to equate code with language, and focus their analyses on either functional distributions, or the definition of situations. critics have pointed out that blom and gumperz (1972) provide scant detail of actual language use in their description of the verbal repertoire of hemnesberget. maehlum (1996) is particularly critical of the suggestion that bokmål and ranamål comprise separate codes. she argues that, in other rural areas of norway, local and standard dialects are not nearly as discrete as blom and gumperz suggest. thus, any suggestion that the verbal repertoire of norwegian speakers is comprised by two distinct codes is flawed. maehlum suggests that “local” and “standard” exist not as empirically identifiable, discrete codes, but “as idealized entities: it is their existence as norms which is important” (1996:753, original italics). further, certain phonological or lexical/morphological variables are particularly salient as indicators of particular dialects. this suggests that sociolinguistic variants are available as indexes of various social meanings, but that attempts to define particular codes and the situations in which they occur are problematic. it is perhaps preferable, then, to identify the formal signals of situation or identity available to a group of speakers, and the uses made of these 8 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 9 signals, rather than to assume a priori that dialects, varieties, or languages will be equally salient across groups. more than many subsequent scholars, gumperz seems to have recognized the imperfection of the description of switching as either situational or metaphorical. by 1982, gumperz’s preferred terminology was conversational code switching. (the description and definition of conversational code switching was, however, largely in terms of metaphorical switching.) gumperz acknowledged that it is generally difficult for analysts to identify particular language choices as situational or metaphorical, and that native speakers generally have few intuitions about or recognition of their own conversational code switches. except in cases of diglossia, the association between linguistic form and settings, activities, or participants is highly variable, and rarely definable by static models. since conversational code switching is not amenable to intuitive methods7, and not strictly relatable to macro-sociological categories, gumperz (1982) argued that close analysis of brief spoken exchanges is necessary to identify and describe the function of code switching. on the basis of his analyses of several speech communities, gumperz suggested a list of six code switching functions which “holds across language situations” (75), but is “by no means exhaustive” (81). gumperz suggested quotation marking, addressee specification, interjection, reiteration, message qualification, and “personalization versus objectivization”8 (80) as common functions of conversational code switching. it is noteworthy that the functions of code switching that gumperz identifies are quite similar to the contextualization cues he describes elsewhere in the volume.9 code switching signals contextual information equivalent to what in monolingual settings is conveyed through prosody or other syntactic or lexical processes. it generates the presuppositions in terms of which the content of what is said is decoded. [gumperz 1982:98] like other contextualization cues, language alternation may provide a means for speakers to signal how utterances are to be interpreted—i.e. provide information beyond referential content. 7 gumperz (1982) points out that both subjects in hemnesberget and spanish-english bilinguals in the united states denied any alternation of linguistic form, but even after listening to recordings of themselves and “promising” to refrain from switching, persisted in code switching. 8 the category of “personalization versus objectivization” is somewhat fuzzy, but relates to illocutionary force, evidentiality, and speaker positioning. 9 gumperz, it may be said, makes the comparison in reverse. his discussion of contextualization conventions (gumperz 1982, chapter 6) says that they are “meaningful in the same sense that… the metaphorical code switching of chapter 4 [is] meaningful” (139). 9 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 10 gumperz’s list of code switching functions inspired many subsequent scholars to refine or propose their own lists of functions (e.g. mcclure and mcclure 1988; romaine 1989; nishimura 1997; zentella 1997). however, as auer (1995) suggests, the functions suggested by such lists are often ill defined. the oft-cited category of reiteration, for example, fails to define exactly what is repeated, or why. lists also tend to combine linguistic structures (such as interjection) and pragmatic or conversational functions (message qualification, addressee specification) without attempting to trace the relationship between forms and functions. although such lists may provide a useful step in the understanding of conversational code switching, they are far from a satisfactory answer to the questions of why switching occurs as it does and what functions it serves in conversation. noting a number of studies that have, following gumperz (1982) suggested similar taxonomies of functions, bailey (2002) notes, “the ease with which such categories can be created – and discrepancies between the code switching taxonomies at which researchers have arrived – hint at the epistemological problems of such taxonomies” (77). code switching may serve any of a number of functions in a particular interaction, and a single turn at talk will likely have multiple effects. therefore, any finite list of functions will be more or less arbitrary. again, the suggestion is that it will be preferable to observe actual interaction, rather than starting from assumptions about the general effects of code switching. 2. sociocultural studies of code switching code switching scholarship within sociocultural linguistics may be divided into several (sometimes overlapping) streams. for the purposes of this paper, three broad areas will be discussed: the social psychological approach of myers-scotton’s markedness model (1983, 1993, 1998) and related work; analyses of identity and code choice; and studies of the effect of code switching on talk in interaction. this last category, largely based on conversation analysis, tends to view code switching behavior both as a method of organizing conversational exchange and as a way to make knowledge of the wider context in which conversation takes place relevant to an ongoing interaction. since this wider knowledge is usually analyzable at least partially in terms of identity, the separation between what i here call “interaction and code switching” versus “identity and code switching” is neither absolute nor unambiguous. indeed, the three-part division suggested here should be seen as one of analytic convenience, rather than significant theoretical import. 10 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 11 2.1. myers-scotton’s markedness carol myers-scotton described her markedness model in the book social motivations for codeswitching: evidence from africa (1993).10 according to myers-scotton, each language in a multilingual community is associated with particular social roles, which she calls rights-and-obligations (ro) sets (84). by speaking a particular language, a participant signals her understanding of the current situation, and particularly her relevant role within the context. by using more than one language, speakers may initiate negotiation over relevant social roles. myers-scotton assumes that speakers must share, at least to some extent, an understanding of the social meanings of each available code. if no such norms existed, interlocutors would have no basis for understanding the significance of particular code choices. the markedness model is stated in the form of a principle and three maxims. the negotiation principle, modeled on grice’s (1975) cooperative principle, presents the theory’s central claim. choose the form of your conversational contribution such that it indexes the set of rights and obligations which you wish to be in force between the speaker and addressee for the current exchange. [myers-scotton 1993:113, original italics] three maxims follow from this principle. the unmarked choice maxim directs, “make your code choice the unmarked index of the unmarked ro set in talk exchanges when you wish to establish or affirm that ro set” (114). the marked choice maxim directs, “make a marked code choice…when you wish to establish a new ro set as unmarked for the current exchange” (131). the exploratory choice maxim states, “when an unmarked choice is not clear, use cs [code switching] to make alternate exploratory choices as candidates for an unmarked choice and thereby as an index of an ro set which you favor” (142). thus, the social meanings of language (code) choice, as well as the causes of alternation, are defined entirely in terms of participant rights and obligations. some critics of the markedness model argue that it relies too heavily on external knowledge, including assumptions about what speakers understand and believe. auer (1998) argues that it is possible to account for code switching behavior without appeal to the “conversation-external knowledge about language use” (10) required by the markedness model. of course, it is possible for the 10 myers-scotton discussed similar issues and developed the markedness model in code choice prior to the publication of this book (e.g. myers-scotton 1972, 1976, 1983). myers-scotton 1983 actually laid out the negotiation principle and six maxims, including the unmarked choice and exploratory choice maxims that figure in the refined model. however, as the fullest expression of the model, it is myers-scotton 1993 that has influenced much subsequent work. 11 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 12 analyst to learn which languages are typically used in particular situations via, for example, ethnographic observation. furthermore, one can argue that speakers learn these norms as part of the language socialization process. a stronger criticism remains, however: the markedness model requires the analyst to make assumptions about each individual speaker’s knowledge and understanding of the speech situation. code switching is then explained on the basis of the analyst’s assumptions about speakers’ internal states (including shared judgments about rights and obligations) rather than its effects on the conversation at hand. further, auer (1995) points out that empirical studies have failed to reveal the strong correlations between particular languages and speech activities that the markedness model predicts. nevertheless, the markedness model is probably the most influential and most fully developed model of code switching motivations. myers-scotton continues to refine the model in ways that are consistent with current research on contact linguistics (myers-scotton 1998; myers-scotton and bolonyai 2001) and the so-called standard theory (chomsky 1965) of linguistics (myers-scotton and jake 2001; jake, myers-scotton and gross 2002). 2.2. identity and code switching whereas the markedness model and subsequent work seeks to provide a systematic and generalizable account of the process of code switching, much work in linguistic anthropology, sociolinguistics, and other areas of sociocultural linguistics provide interpretive and interactional understandings of code switching in particular contexts. although this school of sociocultural linguistics has produced its share of broad theoretical work (e.g. milroy and muysken 1995; alvarez-cáccamo 1998, 2000; woolard 2004), it is generally more closely tied to the observation of behavior in particular settings than to generally applicable explanations of linguistic capability. such studies stand as illustrations of the place of code switching in particular social and historical settings, rather than as models for a universal practice or potential (heller 1992). monica heller’s ethnographic observations and sociolinguistic study in quebec and ontario have led her to consider the economics of bilingualism11, and to view code switching as a political strategy (heller 1988b, 1992, 1995, 1999). since languages tend to become associated with idealized situations and groups of speakers, the use of multiple languages “permits people to say and do, indeed to be two or more things where normally a choice is expected” (heller 1988b:93). this strategic ambiguity allows anglophones in quebec, for example, to achieve a position in francophone controlled corporate culture, while still laying claim to an 11 nor is heller unique in brining such an economic perspective to discourse strategies. compare gal (1979, 1988), woolard (1985), hill (1985), et alia. 12 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 13 anglophone identity, with its associated value on the international market. by uniting bourdieu’s (1977) concept of symbolic capital with gumperz’s (1982) discussion of verbal repertoires, heller (1992, 1995) argues that dominant groups rely on norms of language choice to maintain symbolic domination, while subordinate groups may use code switching to resist or redefine the value of symbolic resources in the linguistic marketplace. while heller and others describe the relationship between language and identity in economic or class terms, many scholars have focused on social categories such as ethnicity. rampton’s (1995) work on crossing, a type of code switching practiced by speakers across boundaries of ethnicity, race, or language ‘community,’12 examines the language behavior of asian, afro-caribbean, and anglo adolescents in ‘ashmead,’ uk. language varieties – creole, panjabi, and stylized asian english – typically associated with an ethnic group, are used by non-members to accomplish complex functions. while rampton does find some of the language-crossing-as-mockery discussed in earlier accounts, crossing in various directions also serves to forge a common adolescent group, to dissociate from parents or elders, and to resist endemic stereotypes. rampton defines crossing in terms of metaphorical switching (blom & gumperz 1972), but in so doing he complicates the notions of situational and metaphorical switching, and of contextualization, considerably. he defines situational switching as language alternation (auer 1984) which accomplishes contextualization (gumperz 1982). rampton reminds us that the boundaries of metaphor are not clear cut (cf. lakoff & johnson 1980); similarly, metaphorical and situational switching cannot be easily delimited. his primary interest, though, is in “figurative” code alternation, a category which, for rampton, is identical to double voicing (bakhtin 1981). unlike situational switching, which rampton argues simply replaces the current situational frame with a new one, crossing adds additional contexts through which an interaction must be interpreted. issues of race, ethnicity, and crossing, as well as economic issues of class and domination are prominent in bailey’s (2001, 2002) work on language and identity among dominican americans. bailey’s work focuses on dominican american youth – young people born in the united states to parents from the dominican republic – living in providence, rhode island. dominican americans, according to bailey (2001, 2002) define their ethnic affiliation as at once nonwhite and non-black. that is to say, while, like their african-american peers, bailey’s subjects view themselves as outside the dominant racial category “white,” they also reject identification with african americans based on 12 code switching or crossing as a means to negotiate or comment on ethnic or racial identities is also seen in the work of nishimura (1992), bucholtz (1999), lo (1999), jaffee (2000), torras and gafaranga (2002), et alia. 13 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 14 phenotype or ancestry. in discourse, this complex identity is indexed by shifting uses of nonstandard dominican spanish, caribbean spanish, african american vernacular english, and other nonstandard english varieties. studies of identity and code switching show that close observation of discourse can yield both empirically and theoretically rich understandings of the functions of language variation in social interaction. by tying observations to particular speakers and social actors, rather than moving too readily to discussions of cultural or linguistic norms, scholars can come to detailed, reliable understandings of the place of language in the construction and transmission of social traditions. 2.3. interaction and code switching close observation of discourse is also a hallmark of interactional linguistics, which seeks to understand “the way in which language figures in everyday interaction and cognition” (ochs, schegloff and thompson 1996:2). these studies tend to be greatly inspired by conversation analysis, as well as functional linguistics and linguistic anthropology. a number of studies under this broad umbrella describe both the place of code switching in the language of turn and sequence and the ways that language alternations, like other contextualization cues, make broader contextual knowledge relevant to an ongoing discourse. auer’s 1984 bilingual conversation presented a pioneering study of interaction and code switching. auer argued that gumperz’s conception of situation is problematic, in that it is defined externally, and from the perspective of the analyst. while auer acknowledged that gumperz’s own uses of situational and metaphorical are less clear-cut that some scholars have taken them to be, he nonetheless disapproved of the distinction. [based on blom & gumperz 1972] one would either have to conclude that (in the situational case) code-switching is without social meaning because it is a necessary consequence of certain situational parameters, or that (in the metaphorical case) it is dependent on an (almost) one-to-one-relationship between language choice and situational parameters which can be purposefully violated. [auer 1984:4] far from pre-existing and determining language choice, auer argues that situation is created by talk in interaction. the form of each speaker’s utterances helps to define the unfolding situation. further, this negotiation itself has social meaning. auer’s analyses of italian migrant children in germany did not find significant correlation between topic and language use. he suggests that code switching is not essentially ‘semantic’ in nature, not derived from the ‘meanings’ of the available languages, but rather is “embedded in the sequential development of the conversation” (1984:93). auer found a great preference for subsequent speakers to maintain the language of the previous turn. language alternation was then available to mark contrast, either to bracket a sequence from the preceding 14 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 15 discourse or to negotiate a common language. auer recommended this procedural analysis of language alternation over individualistic analyses based on introspection, or macro-sociological approaches that define the meaning of potential language choices outside of actual language use. several subsequent studies have examined sequential or interactional functions of language alternation. conversation analysts have suggested that code switching may serve to enhance turn selection (li wei 1998; cromdal 2001) or soften refusals (bani-shoraka 2005; li wei 2005), and is a possible resource to accomplish repair (auer 1995; sebba and wooten 1998) or mark dispreferred13 responses (li wei 1998; bani-shoraka 2005). in addition to these interactional functions, empirical studies have examined how switches in language variety make particular elements of situation, speaker identities, or background relevant to ongoing talk (e.g. li wei 1998, 2002; gafaranga 2001). stroud (1998) criticizes approaches to code switching based too strictly in conversation analysis. he suggests that ca, by proscribing argument from ethnographic or macro-sociological evidence, cannot provide satisfactory analysis of language behavior in non-western settings. stroud observes, “[l]anguage use and patterns of code-switching both structure and are structured by indigenous cultural practices” (1998:322), a suggestion that many sociocultural linguists would probably tend to accept. if analysts then ignore cultural information not visible (to them) within discourse data, their analyses risk missing important elements of function and meaning. stroud maintains, “my argument is that conversational code-switching is so heavily implicated in social life that it cannot really be understood apart from an understanding of social phenomena” (1998:322). this vital understanding is often provided by analysts’ focus on populations that they are themselves a part of; however, it may also be desirable to undertake some broader examination of the social context within which discourse takes place. it seems clear that, in order for observations about the contextualizing functions of language use to have validity and reliability, they should be based on close observation of discourse. at the same time, it should not be assumed that all elements relevant to discourse and social interaction are visible to the analyst, particularly when the analyst is not embedded in the particular social structures he or she is studying. we should remember stroud’s (1998) suggestion that discourse analysis be grounded in an understanding of the society within which communication takes place. the optimal approach to understanding these phenomena would thus seem to include ethnographic observation with close 13 in conversation analysis terms, responses which serve to accomplish the projected action of a previous turn are generally considered preferred, while those that work against such accomplishment are dispreferred. for further explanation, see sacks, schegloff and jefferson 1974 and hutchby and woofit 1998. 15 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 16 analysis of discourse, providing an empirical warrant for any theory of discourse interaction. 3. integrated definitions a great many scholars in sociocultural linguistics use a definition of code switching similar to heller’s: “the use of more than one language in the course of a single communicative episode” (1988a:1). auer and myers-scotton, who largely disagree on how or why code switching occurs, nonetheless sound quite similar in their definitions of the phenomenon. auer (1984:1) refers to “the alternating use of more than one language,” while myers-scotton (1993:vii) mentions “the use of two or more languages in the same conversation.” romaine (1989) cites gumperz as the source of this definition. however, these definitions introduce an element not strictly present in gumperz’s definition: “conversational code switching can be defined as the juxtaposition within the same speech exchange of passages of speech belonging to two different grammatical systems or subsystems” (gumperz 1982:59). note that gumperz’s original definition refers to “grammatical systems or subsystems,” while the subsequent restatements refer to languages. while the former is scarcely more concrete or less ambiguous than the latter, it need not be assumed that the two terms are identical. the plural languages seems to suggest discrete varieties (as english, spanish, kiswahili, etc.), while the more equivocal “systems or subsystems” might equally imply languages or elements of a language, such as lexical items, syntactic constructions, and prosodic phenomena. this list of grammatical subsystems is very similar to goffman’s (1979) list of footing cues and virtually identical to gumperz’s (1982) preliminary list of contextualization cues. the attempt to define language and languages is a perennial controversy in linguistics. by defining code simply as a language (or variety of language) without first defining these basic terms, scholars have essentially put off what should be a foundational question. alvarez-cáccamo (1990, 1998, 2000) provides exceptional attempts to define code and code switching. his discussion relies in turn on work by jakobson (1971b; jakobson, fant and halle 1952, inter alia) and gumperz (1982, 1992, inter alia). alvarez-cáccamo (1998) points out that for jakobson, an early adopter of the term code switching who was influenced by information theory, languages have codes; they do not comprise codes. a language user thus makes use of a code or codes when speaking, listening, etc. the precise nature of any language user’s codes cannot be ascertained by an analyst nor by fellow speakers. internal individual codes (senders’ and receivers’) must necessarily differ, as they belong to different minds. but all human minds are also uniquely alike: they produce language and communication, which are formidably universal. therefore, the question whether each person possesses “different”... codes is parallel to the question 16 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 17 whether speakers of the “same” language share a grammar, or whether culture, ideology, etc., is also shared. there are no absolute answers to this, only a pragmatic one: does communication between two persons sufficiently work? [alvarezcáccamo, personal communication] speakers use communicative codes in their attempts (linguistic or paralinguistic) to communicate with other language users. listeners use their own codes to make sense of the communicative contributions of those they interact with. listeners may need to shift their expectations to come to a useful understanding of speakers’ intentions. similarly, speakers may switch the form of their contributions in order to signal a change in situation, shifting relevance of social roles, or alternate ways of understanding a conversational contribution. in other words, switching codes is a means by which language users may contextualize communication. a useful definition of code switching for sociocultural linguistic analysis should recognize it as an alternation in the form of communication that signals a context in which the linguistic contribution can be understood. the ‘context’ so signaled may be very local (such as the end of a turn at talk), very general (such as positioning vis-à-vis some macro-sociological category), or anywhere in between. furthermore, it is important to recognize that this signaling is accomplished by the action of participants in a particular interaction. that is to say, it is not necessary or desirable to spell out the meaning of particular code switching behavior a priori. rather, code switching is accomplished by parties in interaction, and the meaning of their behavior emerges from the interaction. this is not to say that the use of particular linguistic forms has no meaning, and that speakers “make it up as they go.” individuals remember and can call on past experiences of discourse. these memories form part of a language user’s understanding of discourse functions. therefore, within a particular setting certain forms may come to recur frequently. nonetheless, it is less interesting (for the current author at least, and probably for the ends of sociocultural linguistic analysis) to track the frequency or regularity of particular recurrences than to understand the effect of linguistic form on discourse practice and emergent social meanings. to recapitulate, then, code switching is a practice of parties in discourse to signal changes in context by using alternate grammatical systems or subsystems, or codes. the mental representation of these codes cannot be directly observed, either by analysts or by parties in interaction. rather, the analyst must observe discourse itself, and recover the salience of a linguistic form as code from its effect on discourse interaction. the approach described here understands code switching as the practice of individuals in particular discourse settings. therefore, it cannot specify broad functions of language alternation, nor define the exact nature of any code prior to interaction. codes emerge from interaction, and become relevant when parties to discourse treat them as such. 17 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 18 references alvarez-cáccamo, celso. 1990. “rethinking conversational code-switching: codes, speech varieties, and contextualization.” in kira hall, jean-pierre koenig, michael meacham, sondra reinman and laurel sutton (eds.) proceedings of the sixteenth annual meeting of the berkeley linguistics society, 3-16. berkeley: berkeley linguistics society. --. 1998. “from ‘switching code’ to ‘code-switching’: towards a reconceptualization of communicative codes.” in peter auer (ed.) codeswitching in conversation: language, interaction, and identity, 29-48. london: routledge. --. 2000. “para um modelo do ‘code-switching’ e a alternancia de variedades como fenomenos distintos: dados do discurso galego-portuges/espanhol na galiza” (toward a model of ‘code-switching’ and the alternation of varieties as distinct phenomena: data from galician portuguese/spanish discourse in galicia). estudios de sociolinguistica 1(1): 111-128. auer, peter. 1984. bilingual conversation. amsterdam: john benjamins. --. 1995. “the pragmatics of code-switching: a sequential approach.” in lesley milroy and pieter muysken (eds.) one speaker, two languages: crossdisciplinary perspectives on code-switching, 115-135. cambridge: cambridge university press. --. 1998. code-switching in conversation: language, interaction, and identity. london: routledge. azuma, shoji. 1991. “two level processing hypothesis in speech production: evidence from intrasentential code-switching.” papers from the regional meetings, chicago linguistic society, 27(1): 16-30. bailey, benjamin. 2001. “the language of multiple identities among dominican americans.” journal of linguistic anthropology 10(2): 190-22 --. 2002. language, race, and negotiation of identity: a study of dominican americans. new york: lfb scholarly publishing. bakhtin, mikhail. 1981. the dialogic imagination. caryl emerson and michael holquist (trans.). austin: university of texas press. bani-shoraka, helena. 2005. language choice and code-switching in the azerbaijani community in tehran: a conversation analytic approach to bilingual practices. uppsala, sweden: acta universitatis upsaliensis. barker, george. 1947. “social functions of language in a mexican-american community.” acta americana 5: 185-202. belazi, heidi, edward rubin, and almeida jacqueline toribio. 1994. “code switching and x-bar theory: the functional head constraint.” linguistic inquiry 25(2): 221-237. benson, erica. 2001. “the neglected early history of codeswitching research in the united states.” language & communication 21: 23-36. blom, jan-petter, and john gumperz. 1972. “social meaning in linguistic structures: code switching in northern norway.” in: john gumperz and del 18 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 19 hymes (eds.): directions in sociolinguistics: the ethnography of communication, 407-434. new york: holt, rinehart, and winston. bordieu, pierre. 1977. “the economics of linguistic exhanges.” social science information 16, 645-668. bucholtz, mary. 1999. “you da man: narrating the racial other in the production of white masculinity.” journal of sociolinguistics 3(4): 443-460. bucholtz, mary, and kira hall. 2005. “identity and interaction: a sociocultural linguistic approach.” discourse studies 7(4-5). cenoz, jasone and fred genesee. 2001. trends in bilingual acquisition. amsterdam: john benjamins. chomsky, noam. 1965. aspects of the theory of syntax. cambridge, ma: mit press. cromdal, jakob. 2001. “overlap in bilingual play: some implications of codeswitching for overlap resolution.” research on language and social interaction 34(4): 421-451. dil, anwar s. 1971. “introduction.” in john j. gumperz, language in social groups: essays by john j. gumperz. stanford: stanford university press. disciullo, anna-maria, and edwin williams. 1987. on the definition of a word. cambridge, ma: mit press. ervin-tripp, susan. 1964. “an analysis of the interaction of language, topic and listener.” american anthropologist 66(6): part 2, 86-102. fano, robert m. 1950. “the information theory point of view in speech communication.” journal of the acoustical society of america 22, 691-696. ferguson, charles. 1959. “diglossia.” word 15, 325-340. fishman, joshua. 1967. “bilingualism with and without diglossia; diglossia with and without bilingualism.” journal of social issues 23(2): 29-38. fotos, sandra. 2001. “codeswitching by japan’s unrecognized bilinguals: japanese university students’ use of their native language as a learning strategy.” in mary goebel noguchi and sandra fotos (eds.) studies in japanese bilingualism. clevedon: multilingual matters. friedrich, paul. 1972. “social context and semantic feature: the russian pronominal usage.” in john gumperz and dell hymes (eds.) directions in sociolinguistics: the ethnography of communication, 270-300. new york: holt, rinehart and winston. gafaranga, joseph. 2001. “linguistic identities in talk-in-interaction: order in bilingual conversation.” journal of pragmatics 33(12): 1901-1925. gal, susan. 1979. language shift: social determinants of linguistic change in bilingual austria. new york: academic press. --. 1988. “the political economy of code choice.” in monica heller (ed.) codeswitching: anthropological and sociolinguistic perspectives, 243-261. berlin: mouton de gruyter. grice, h. paul. 1975. “logic and conversation.” in peter cole and jerry l. morgan (eds.): speech acts, 41-55. new york: academic press. goffman, erving. 1979. “footing.” semiotica 25, 1-29. 19 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 20 --. 1981. forms of talk. philadelphia: university of pennsylvania press. gumperz, john. 1958. “dialect differences and social stratification in a north indian village.” american anthropologist 60, 668-681. --. 1961. “speech variation and the study of indian civilization.” american anthropologist 63, 976-988. --. 1964a. “hindi-punjabi code-switching in delhi.” in h. hunt (ed.): proceedings of the ninth international congress of linguistics, 1115-1124. the hague: mouton. --. 1964b. “linguistic and social interaction in two communities.” american anthropologist 66(6): part 2, 137-153. --. 1982. discourse strategies. cambridge: cambridge university press. --. 1992. “contextualization revisisted.” in peter auer and aldo di luzo (eds.) the contextualization of language, 39-53. amsterdam: john benjamins. halmari, helena. 1997. government and codeswitching: explaining american finnish. amsterdam: john benjamins. heller, monica. 1988a. codeswitching: anthropological and sociolinguistic perspectives. berlin: mouton de gruyter. --. 1988b. “strategic ambiguity: code-switching in the mangagement of conflict.” in monica heller (ed.) codeswitching: anthropological and sociolinguistic perspectives, 77-96. berlin: mouton de gruyter. --. 1992. “the politics of codeswitching and language choice.” in carol eastman (ed.) codeswitching, 123-142. clevedon: multilingual matters. --. 1995. “code-switching and the politics of language.” in lesley milroy and pieter muysken (eds.) one speaker, two languages: cross-disciplinary perspectives on code-switching. cambridge: cambridge university press. --. 1999. linguistic minorities and modernity: a sociolinguistic ethnography. london: longman. hill, jane. 1985. “the grammar of consciousness and the consciousness of grammar.” american ethnologist 12(4): 725-737. hutchby, ian and robin woofit. 1998. conversation analysis: principles, practices, and applications. malden, ma: polity press. hymes, dell. 1964. “introduction: toward ethnographies of communication.” american anthropologist 66(6): part 2, 1-34. jaffe, alexandra. 2000. “comic performance and the articulation of hybrid identity.” pragmatics 10(1): 39-59. jake, janice, carol myers-scotton and steven gross. 2002. “making a minimalist approach to codeswitching work: adding the matrix language.” bilingualism: language and cognition 5(1): 69-91. jakobson, roman. 1971a. “results of a joint conference of anthropologists and linguists.” in selected writings, volume ii, 554-567. the hague: mouton. --. 1971b. “linguistics and communication theory.” in selected writings, volume ii, 570-579. the hague: mouton. 20 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 “code switching” in sociocultural linguistics 21 jakobson, roman, gunnar fant and morris halle. 1952. preliminaries to speech analysis: the distinctive features and their corelates. cambridge, ma: mit press. jakobson, roman and morris halle. 1956. fundamentals of language. the hague: mouton. joshi, aravind. 1985. “how much context-sensitivity is necessary for assigning structural descriptions: tree adjoining grammars.” in d. dowty, l. karttunen, and a. zwicky (eds.) natural language parsing. cambridge: cambridge university press. lakoff, george and mark johnson. 1980. metaphors we live by. chicago: university of chicago press. li wei. 1998. “the ‘why’ and ‘how’ questions in the analysis of conversational code-switching.” in peter auer (ed.): code-switching in conversation: language, interaction and identity, 156-176. london: routledge. --. 2005. “‘how can you tell?’ toward a common sense explanation of conversational code-switching.” journal of pragmatics 37(3): 375-389. lo, adrienne. 1999. “codeswitching, speech community membership, and the construction of ethnic identity.” journal of sociolinguistics 3-4, 461-479. macswann, jeff. 2000. “the architecture of the bilingual language faculty: evidence from intrasentential code switching.” bilingualism: language and cognition 3(1): 37-54. maehlum, brit. 1996. “codeswitching in hemnesberget – myth or reality?” journal of pragmatics 25, 749-761. mcclure, erica and malcolm mcclure. 1988. “macroand micro-sociolinguistic dimensions of code-switching in vingard (romania).” in monica heller (ed.) codeswitching: anthropological and sociolinguistic perspectives, 25-51. berlin: walter de gruyter. milroy, lesley and pieter muysken. 1995. one speaker, two languages: crossdisciplinary perspectives on code-switching. cambridge: cambridge university press. myers-scotton, carol. 1972. choosing a lingua franca in an african capital. edmonton: linguistic research. --. 1976. “strategies of neutrality: language choice in uncertain situations.” language 52(4): 919-941. --. 1983. “the negotiation of identities in conversation: a theory of markedness and code choice.” international journal of the sociology of language 44, 115-136. --. 1993. social motivations for codeswitching: evidence from africa. oxford: clarendon press. --. 1998. codes and consequences: choosing linguistic varieties. new york: oxford university press. myers-scotton, carol and agnes bolonyai. 2001. “calculating speakers: codeswitching in a rational choice model.” language in society 30, 1-28. 21 nilep: “code switching” in sociocultural linguistics published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 22 myers-scotton, carol and janice jake. 2001. “four types of morpheme: evidence from aphasia, code switching, and second-language acquisition.” linguistics 38(6): 1053-1100. nishimura, miwa. 1992. “language choice and in-group identity among canadian niseis.” journal of asian pacific communication 3(1): 97-113. --. 1997. japanese-english code-switching: syntax and pragmatics. new york: p. lang. ochs, elinor, emanuel schegloff and sandra thompson. 1996. interaction and grammar. cambridge: cambridge university press. poplack, shana. 1980. “sometimes i'll start a sentence in spanish y termino en espanol: toward a typology of code-switching.” linguistics 18(233-234): 581-618. rampton, ben. 1995. crossing: language and ethnicity among adolscents. london: longman. romaine, suzanne. 1989. bilingualism. oxford: basil blackwell. sacks, harvey, emanuel schegloff and gail jefferson. 1974. “a simplest systematics for the organization of turn taking for conversation.” language 50, 696-735. sankoff, david, and shana poplack. 1981. “a formal grammar for codeswitching.” papers in linguistics 14(1-4): 3-45. sapir, edward. 1929. “the status of linguistics as a science.” language 5(4): 207-214. sebba, mark and tony wooten. 1998. “we, they and identity: sequential versus identity-related explanation in code-switching.” in peter auer (ed.): codeswitching in conversation: language, interaction and identity, 262-286. london: routledge. stroud, christopher. 1998. “prespectives on cultural variability of discourse and some implications for code-switching.” in peter auer (ed.): code-switching in conversation: language, interaction and identity, 321-348. london: routledge. torras, maria-carme and joseph gafaranga. 2002. “social identities and language alternation in non-formal institutional bilingual talk: trilingual service encounters in barcelona.” language in society 31(4): 527-548. vogt, hans. 1954. “language contacts.” word 10(2-3): 365-374. weinreich, uriel. 1953. languages in contact. the hague: mouton. woolard, katherine. 1985. “language variation and cultural hegemony.” american ethnologist 12(4): 738-748. --. 2004. “codeswitching.” in alessandro duranti (ed.) a companion to linguistic anthropology, 73-94. malden, ma: blackwell. zentella, ana celia. 1997. growing up bilingual. malden, ma: blackwell. 22 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/1 doi: https://doi.org/10.25810/hnq4-jv62 colorado research in linguistics 6-2006 “code switching” in sociocultural linguistics chad nilep recommended citation microsoft word !nilep_cril_2006_final.doc microsoft word olmsted-cril2021-proof_final.docx 1 how to do things with memes: creating community through the sociopragmatics of star wars prequel memes carolyn olmsted university of colorado boulder in the 1950s and 1960s, j. l. austin changed the face of pragmatic linguistic analysis with his work how to do things with words. famously claiming that language does not simply describe the world but also changes it, austin (1962) established the theory of speech acts. this paper builds on austin’s research by investigating internet memes as a kind of speech act; hence the title “how to do things with memes.” focusing in particular on a specific genre of memes that incorporates images and discourse from the american space epic media franchise star wars, the paper explores the following question: what are the sociopragmatic functions of memes used on the internet today by younger generations of internet users? the conversation analytic concept of adjacency pairs is used to understand the redistribution, recontextualization, and remediation of original media sources into memes, together with linguistic anthropological research that interrogates recontextualization as a kind of performance (baumann and briggs 1990). an investigation of the illocutionary forces behind select star wars memes exposes what exactly these memes are “doing” in their respective internet spheres. specifically, the paper outlines how these memes function to build community, as illustrated by sociocultural linguistic work on identity (bucholtz and hall 2004, 2005) and social semiotic concepts such as dual indexicality (hill 1995). through processes of memetic participation, digital users build community and draw closer together by expanding upon their existing communicative repertoires (rymes 2012, 2014). keywords: sociolinguistics, memes, internet linguistics, pragmatics 1. introduction in his famous work how to do things with words, j. l. austin (1962) changed the face of modern linguistic analysis. austin described language not simply as a tool to convey information, but as a way to change the world around us via speech acts. speech acts can accomplish many different things: canonical speech acts include a judge sentencing a man to death or a priest marrying a happy couple, but smaller speech acts can also be performed in everyday conversations. the speech act need not even be as direct; that is, saying “it’s stuffy in here” may prompt someone to open a window, thus changing the world. while these words are not expressed in a syntactic form that we normally associate with a request, their pragmatic meaning—or illocutionary force, as austin calls it—carries the potential to change the world. this paper suggests that everyday actions such as shaping a community can be accomplished with words and the illocutionary forces colorado research in linguistics, volume 25 (2021) 2 they carry. the analysis applies austin’s understanding of the world-changing role played by speech acts to the digital genre of memes, focusing in particular on a specific genre of memes that incorporates images and discourse from the american space epic media franchise star wars. using this genre of memes, known to star wars fans as “prequel memes”, i explore the following question: what are the sociopragmatic functions of memes used on the internet today? this paper is divided into two parts. in the first section, titled “the sociopragmatics of memes,” i explain how memes work through the processes of redistribution, recontextualization, and remediation. these concepts are taken from the linguistic field known as computer-mediated discourse (hereafter cmd). much early cmd research focused on text-based synchronous language used in formats such as chat rooms or text-based asynchronous language used in formats such as email. my paper, however, focuses on the multimodal genre of memes specifically. i apply the above concepts to understand the sociopragmatic work that memes are doing as they circulate across digital users, especially their community-building properties, rather than simply describing their form and distribution. in the second section, titled “star wars prequel fans as a community of practice,” i offer a case study on the star wars prequel meme community. star wars prequel fans, as i call them throughout this paper, make heavy use of memes to define who is and is not considered a part of the community. this type of identity work is discussed by mary bucholtz and kira hall (2004, 2005) as involving the processes of adequation and distinction, two concepts that figure prominently in my analysis. one of the most popular memes used in this community is what i refer to as the “general kenobi” meme, explained in section 3.1, which is based on an adjacency pair from the star wars prequel films. past research in the field of conversation analysis has tended to focus on universal aspects of adjacency pairs in conversational turn-taking, but i focus instead on how community members use them to display and participate in community building and belonging. i conclude with an analysis of how, exactly, star wars prequel fans are doing things with memes. my analysis is informed by sociolinguistic work on the ways that identity emerges through the telling of formulaic jokes (hall 2019), stancetaking (bucholtz, et al. 2011), adequation and distinction (bucholtz & hall 2004), and dual indexicality (hill 1995). in this sense, what takes place in meme-sharing among star wars prequel fans is closely aligned with the sociolinguistic processes associated with identity work more generally. my analysis uncovers these parallels, how to do things with memes 3 while also demonstrating some of the unique affordances offered by memes to the building of communities online. 2. the sociopragmatics of memes memes are one of the most widespread forms of communication on the internet today. the word meme has typically been applied to digital media to refer to “something that’s remade and recombined, spreading as an atom of internet culture” (mcculloch 2019); however, there is a general lack of consensus among users online about what exactly a meme is. stripped down to their basic parts, it seems that memes are jokes spread online, usually involving images, and almost always involving some type of language. when we look at traditional mediums of language through the lens of speech act theory, we can see, to borrow a phrase from austin, “how to do things with words”. similarly, if we look at computer-mediated discourse (hereafter cmd) through the same lens, we can see “how to do things with memes”. through processes known to cmd researchers as redistribution, recontextualization, and remediation, memes carry a kind of illocutionary force that builds and strengthens community. 2.1. redistribution the first and most important part of how a meme gains its illocutionary force is through redistribution. redistribution, which involves simply sending a message or form of media unaltered to someone else, is how a meme makes its way through and around communities. without redistribution, there would be no memes at all; memes take their meaning and illocutionary force only from being shared within communities. redistribution, being the most fundamental component of memes, is also a process central to non-digital forms of communication. what may be seen as the predecessors of memes, and are described as such by mcculloch (2019), belong to the pre-internet era and relied solely on physical circulation among people. one example is that of newspaper clippings; to share a newspaper article with another, one must cut it out and give it or send it to them physically. an early form of digital communication, still prominent today, that overtly relies on redistribution is that of chain emails, a genre continuous with earlier non-digital “chain letters” sent by mail. chain emails often have a similar format to that of the one in figure 1. colorado research in linguistics, volume 25 (2021) 4 figure 1. example of a chain email (hgrant 2012) chain emails like this one are designed to be sent, or redistributed, to a large group of people unaltered. the concept of redistribution is therefore inherently simple in that no part of the original message is changed; however, cmd theorists would argue that the act of sending this email into new contexts potentially invites a change of social meaning that is worthy of analysis. this is because of their reach and targeted audience. while redistribution has certainly been around for as long as people have been sharing information, cmd is a rather new way of accomplishing this. cmd allows vast amounts of information to be distributed among exponentially more people than ever before in very short amounts of time. because of this, memes can very quickly bring together a large community of people. additionally, redistribution can select the targeted audience for the redistributed content. in figure 1, the chain email is directed at those who both would be flattered by the compliment “beautiful” and who would like to share that compliment with others. the act of redistribution and those who perform it can have any kind of target audience in mind: family, friends, even star wars prequel fans. it is important to recognize that the concept of redistribution is foundational to the concepts of recontextualization and remediation, which more overtly focus on the ways texts take on new meanings as they undergo circulation. 2.2. recontextualization the concept of recontextualization relies heavily on the idea of a communicative repertoire. betsy rymes (2012:216) describes a communicative repertoire as “the collection of ways individuals use language and other means of communication… to function effectively in the how to do things with memes 5 multiple communities in which they participate” and as composed of “mass-mediated cultural elements, circulated, often, via viral internet sources”. to rymes, the most important observation is that the “repertoire elements are… catchy, memorable, or dramatic”, which makes them “highly recontextualizable bits”. to say that repertoire entries are recontextualizable means that they can easily be removed from their original context and placed in a different social context, a process linguistic anthropologists bauman and briggs (1990) identify as entextualization. recontextualization, which follows decontextualization, is an expansion of redistribution in that the text is likewise redistributed and entered into circulation; however, the concept focuses not so much on the act of redistribution but rather on the transformation that occurs when the text is removed from one context and placed into another. essentially, cmd researchers are interested in how and why a text may become highly quotable and easily recognized as it enters new contexts. this concept can be applied to many memes originating from various works of media such as star wars or lord of the rings as well as viral videos, as seen below. figure 2. “cedar rapids” meme (illusions 2018) figure 3. “farming” meme (edwards 2016) figure 2 is a still from a viral video featuring hillary clinton in which she says, “i’m just chillin’ here in cedar rapids”, and figure 3 is a quote from the star wars film rogue one. these images are often posted around the internet with no additional text or other images, which redistributes as well as recontextualizes them. because they are not simply the original work presented in an original context but rather an explicit reference to an earlier piece of media (a process i will discuss colorado research in linguistics, volume 25 (2021) 2 later as remediation), they have advanced from simply featuring redistribution to featuring recontextualization as well. additionally, the references’ widespread use among many digital users in a given community, such as clinton supporters or star wars prequel fans, shows that they have become part of the communicative repertoires of their respective communities. these types of memes are extremely typical in groups focused on a single franchise or piece of media. the memes seen above are only one example of the vast collection of recontextualizable quotes used in their respective internet domains. however, not all the memes circulating in these groups feature only recontextualization. as suggested earlier, they can also be combined with other repertoire elements in a phenomenon known as remediation. 2.3. remediation as described by bolter and grusin (1999), the term remediation refers to the way that digital media is constantly recalling or incorporating its media predecessors—most notably, “older” forms of media such as film, television, photography, or even handwriting. when analyzing memes, cmd theorists often use the term remediation to reference how one meme may incorporate several different works of media. these can be “older media” fictional works like television shows or movies, but they can also involve other repertoire elements known as meme templates. meme templates are images with blank slots for users to fill in their own text or images, as seen in figures 6 and 8; there are even “meme generator” websites that enable users to produce new memes based on popular, already circulating templates. these templates are thus extremely recontextualizable and widely circulated, perhaps even more so than other repertoire elements. in my research on memes, i have found that most examples of remediation can be classified in two ways: normative remediation between different works of media, and combinatorial remediations that alter and “remix” two media pieces into a novel form. this last form, in particular, is characterized by the heavy use of meme templates, often redesigned for use in a specialized knowledge space. the “form” of the meme is retained even though the material within it changes, similarly to how syntax functions in a sentence. the first type, remediation between different works of media, is probably the most easily identifiable type of remediation. it involves explicit references to multiple works, with several unaltered repertoire elements from each inserted into, or perhaps more accurately, on top of each other. the unaltered repertoire elements are a hallmark of this first type of remediation; if they how to do things with memes 3 were altered, they would be of the second type. the first type may be described as cutting and pasting, both digitally and physically. the content creator simply cuts a face, character, object, or other recognizable repertoire element out of one image and enters it into a scene, background, or context from another one. this type can be seen in the examples below. figure 4. example of remediation between star wars and the office (u/poopypants1234321 2017) figure 5. example of remediation between lord of the rings and monty python and the holy grail (u/futurarmy 2019) in these examples, figure 4 involves a remediation of a scene from star wars and a character from the office and figure 5 is a remediation of a scene from lord of the rings and a scene from monty python and the holy grail. to understand any of these memes, or find them funny, the reader must be familiar with both works of media referenced in the meme and have the characters, scenes, or other references in the meme in their own communicative repertoire. for example, if a reader were familiar with lord of the rings but did not have monty python and the holy grail in their communicative repertoire, they would be confused and not find the meme very funny, if at all. however, someone familiar with the references would understand that in the scene referenced, the man pictured (boromir of lord of the rings) is holding and almost succumbing to the evil of the colorado research in linguistics, volume 25 (2021) 2 one ring, a powerful object of world-destroying proportions. juxtaposing this serious scene with a reference from the very comedic film monty python and the holy grail turns the seriousness on its head for comedic effect. one unfamiliar with one of these references would not understand the juxtaposition in mood and therefore miss the humor. even though this type of remediation requires broader forms of specialized knowledge, it can be extremely popular if both references are something many people have in their communicative repertoires. the second type of remediation is the type i find the most interesting: remediation that goes even further than the first type and “remixes” the two original pieces into something distinctively new. this type relies heavily on communicative repertoire entries such as meme templates and creates new content by using an existing form as a vehicle for understanding. the “syntax” of the meme is retained, but the elements are replaced with scenes requiring specialized knowledge. instead of taking two separate images and cutting-and-pasting them together, as in the first two types, this type creates a new image from a blank template. of course, the new image is necessarily evocative of the element or elements appearing in the original template; otherwise, the reference would be unnoticeable. however, no elements in the original image, other than the template itself, appear in the new image. usually, this form manifests itself in the form of a new meme template made using only references to a single work of media. in the examples below, i include the original meme template as well as the remediated form. figure 6. “drake” meme template (n.a. 2019 “drake hotline bling meme generator”) figure 7. remediation of “drake” meme template and kermit the frog (w__a__c 2019) how to do things with memes 1 the meme in figure 7 is a remediated form of the “drake” meme seen in figure 6, in which drake, the man in the orange jacket, first expresses dislike for something (upper image), and then preference for something (lower image). meme creators fill the blank boxes next to these two images with texts or pictures representing their ideas or opinions. a remediated example of the drake meme can be seen in figure 7, which shows kermit the frog in stances that recall the original. in this meme, the creator expresses preference for using kermit’s picture instead of drake’s “because he is cute”. if a person viewing the meme is unfamiliar with the drake meme format, they will not understand the first panel’s intertextuality with the drake meme and might not understand the concept that the creator is trying to express. this type of remediation relies the heaviest on communicative repertoire, as shown by the fact that they are difficult to understand fully without knowledge of the meme template they are based on. since they are new, separate images, they often take more effort to create than the other types, due to their type of remediation. rather than simply cutting and pasting images together, these creators replace the original images in a meme template with another set of images that recall the sense of original meme even while advancing something new. because these memes are so much harder to create and understand, the question is raised: why create them at all? they are frequently well-received and rather popular, but only within the communities of practice they target; the people who understand the meme recognize the reward and effort put in. this observation is key to my argument regarding how these memes function to build community. i suggest that many of these memes are created and shared out of a desire to keep communities “pure”; that is, they incorporate characters or concepts from works viewed as central to the communities that exchange them. this ideology can be seen in the meme below, which is also based on the drake meme in figure 6. colorado research in linguistics, volume 25 (2021) 2 figure 8. remediation of “drake” meme template and the star wars prequels (u/avatyler 2017) this meme features a character from the star wars franchise as the replacement for drake in the meme from figure 6 and provides its reasoning for doing so, making it a sort of “meta-meme.” the creator of this meme, who is a member of the subreddit r/prequelmemes, displays a preference for keeping the community “pure” by encouraging members to use memes with content taken only from the star wars prequel films. the accompanying text indicates that community members should replace drake with the star wars character kit fisto (a fan-favorite film character beloved for his smile), shown in the left panels of the meme. this is how this particular type of remediation builds community: by appealing to an internally recognized communicative repertoire. 3. star wars prequel fans as a community of practice this section analyzes how the processes of redistribution, recontextualization, and remediation are deployed within a specific community of practice: digital users who create and exchange star wars prequel memes. prequel memes are a subcategory of star wars memes based on the franchise’s prequel films: the phantom menace, attack of the clones, and revenge of the sith. prequel memes are extremely popular among fans of these films. one of the main online spaces in which prequel memes are published is on the social media website reddit, as we saw above in the remediated drake meme. subreddits resemble forums and are a type of community dedicated to how to do things with memes 3 one topic. the subreddit r/prequelmemes, which is extremely popular with over 1.8 million subscribers, allows only posts with at least tangential relation to the star wars prequels. most prequel memes simply involve quotes from the films. sometimes they are recontextualized into a new joke based on the quote; sometimes the joke is simply the stating of the quote itself. the community is extremely aware of their frequent quoting practices, as seen in this meme of a scene from revenge of the sith. figure 9. an example of a star wars prequel meme that uses a quote with no captions (u/dhogan73 2018) the quote in the original scene from which this picture is taken is “you underestimate my power” (lucas 2005), said by anakin skywalker, the character in the picture, in the middle of the bitter duel at the film’s climax. the joke in this meme, then, is that when outsiders (represented by “them”) say that it is impossible to understand such memes if the quote from the film is not provided in the subtitles, those who are in the community can do just that. that is, for star wars prequel fans, the image alone recalls the quote “you underestimate my power”, providing a perfect response to outsiders who lack the specialized knowledge needed to interpret the image. 3.1. “general kenobi” memes and adjacency pairs an extremely popular and often-quoted prequel meme is what i refer to as the “general kenobi” meme. to understand the “general kenobi” meme, one must first understand the colorado research in linguistics, volume 25 (2021) 4 linguistic phenomenon known as adjacency pairs. adjacency pairs is a concept taken from conversation analysis and refers to a type of conversational turn-taking in which two speakers produce two utterances in succession. the first utterance (or first-pair part) elicits an utterance in response (or second-pair part). second-pair parts are discussed in conversation analytic literature as either preferred or dispreferred (see, e.g., kitzinger and frith 1999), meaning that the second speaker will either respond with the expected answer or type of answer, which would be preferred, or the unexpected answer or type of answer, which would be dispreferred. for example, if a firstpair part were an invitation, the second-pair part could either be an acceptance or a refusal; the acceptance would be preferred and the refusal would be dispreferred. while conversation analysts often discuss preference expectations as part of a general grammar shared by speakers of a language, my work focuses on how members of a particular community—in this case, star wars prequel fans—collaboratively participate in novel adjacency pairs as a display of communal belonging. the “general kenobi” meme is based on an exchange between two characters in the third prequel film, revenge of the sith, seen below in figure 10. figure 10. the “general kenobi” meme (lucas 2005) how to do things with memes 5 in this scene, obi-wan kenobi, the man with the beard, greets his long-time adversary, general grievous, the robot-looking alien. this scene is one of the most popular memes in the r/prequelmemes community; whenever someone on the forum says “hello there!”, there will be many responses of “general kenobi!” 3.2. examples from tinder exchanges the “general kenobi'' exchange is not a normative adjacency pair; however, it has many similarities to both the concept and execution of adjacency pairs. the exchange is often replicated among star wars prequel fans and acts as a gatekeeping device to signal membership in this ingroup. the first-pair part of this adjacency pair is the “hello there!”; the second-pair part preferred response, for these fans, is “general kenobi!”. the preferred response, if given, signals to the first speaker that the second speaker is part of the in-group of star wars prequel fans. however, if the second-pair part is a dispreferred response, the first speaker knows that the second speaker is not a star wars prequel fan, or at least not the kind that memorizes dialogue from the films. an example of a dispreferred response that leads to a joke can be seen in the exchange and accompanying meme in figure 11. figure 11. a dispreferred response to the “general kenobi” first-pair part (u/notfredrhodes 2018) colorado research in linguistics, volume 25 (2021) 6 in the exchange from the dating app tinder, seen on the left, the first messenger offers a “hello there!”, but the second messenger gives a dispreferred response of “heyhey hows your evening been?”, signaling to the first messenger that she does not know the context or expected response to the first-pair part. interestingly, this exchange was posted alongside a recontextualization of the “general kenobi!” meme seen above in the same image file. this recontextualized meme puts the text of the tinder conversation over two frames from the film. this example is illuminated by the social theoretical framework of adequation and distinction, two “tactics of intersubjectivity” described by mary bucholtz and kira hall (2004, 2005) as involved in identity production. according to these authors, adequation “involves the pursuit of socially recognized sameness. in this relation, potentially salient differences are set aside in favor of perceived or asserted similarities that are taken to be more situationally relevant” (2004:383). adequation is not very present in this example, but will clearly be seen in the second example. the second concept is that of distinction. bucholtz and hall write that distinction is “the mechanism whereby salient difference is produced. distinction is therefore the converse of adequation, in that in this relation difference is underscored rather than erased” (2004:384). distinction can very clearly be seen in figure 11. the woman in the tinder conversation gives a dispreferred response, and the first messenger creates a meme mocking her lack of star wars knowledge. by doing this, the meme creator is “establishing a dichotomy,” as hall and bucholtz put it, between himself1 and the woman. by defining her as “not a star wars fan” because she does not get the reference and give the preferred response, he conversely defines “star wars fans” as those who do get the reference and give the preferred response. this exchange provides a strong example of how community is created through distinction, but in figure 12, below, adequation is instead the primary process used. a woman named taylor messages the first-pair part “hello there” and her interlocutor responds with the second-pair part “general kenobi!”. how to do things with memes 7 figure 12. a preferred response to the “general kenobi” first-pair part (u/ashman508 2020) it is difficult to know the exact timing of the exchange due to the asynchronous nature of the image, but it appears that taylor responds while the messenger is typing his follow-up message, because his second message seems to be a post-expansion based on his second-pair part of “general kenobi”, and does not acknowledge taylor’s intervening message of “thank god”. after this brief misunderstanding, the messenger then gives another quote from the prequels “a surprise to be sure, but a welcome one!”, from the prequel film the phantom menace (lucas 1999). he captions his reddit post “this is where the fun begins!”, which is yet another quote from the same film. the more than 200 comments that respond to this post are also very interesting. most provide still more prequel quotes; others make approving statements like “marriage material”, indicating that taylor counts as part of the community. however, one comment says “be careful. there are fakers out there who only do it to snag our poor lads hearts without actually caring about star wars”. bucholtz and hall (2004:384) write that the tactic of distinction “has a tendency to reduce colorado research in linguistics, volume 25 (2021) 8 complex social variability to a single dimension: us versus them”. in this comment, the possibility of “fakers” creates a clear “them” for the community of “our poor lads” to be in opposition with. additionally, the “lads” comment suggests that the community is thought to be made up mostly, if not entirely, of men, which in turn implies that women are the ones most likely to be “fakers”. these “fakers” who do not “actually [care] about star wars” are apparently a menace to the community that need to be carefully surveilled, again showing the importance of adequation and distinction to star wars prequel fan identity construction. 4. conclusions i began this paper with the question: how does one do things with memes? although i have only scratched the surface with the research displayed here, i believe that the answer lies in a type of illocutionary force associated with all the memes i have discussed: community-building. when one person shows a meme to another, they are not simply redistributing it. sharing memes, like sharing any form of humor, brings people closer together. by laughing at a meme together, community members demonstrate that they get the “joke” and thereby foster a sense of belonging. hall (2019:507) describes a comparable phenomenon of formulaic jokes told by urban youth in new delhi, india: “formulaic jokes are massively distributed, yet […] also take on specialized meanings as they enter into localized interactions forged within specific communities”. this closely mirrors the “general kenobi” memes described in this paper; they have a specialized meaning of community-belonging in only the specific community of star wars prequel fans. to outsiders, the “general kenobi” adjacency pair would be seen simply as a movie quote. the specialized meaning only comes to exist through repeated localized interactions within the community. citing bucholtz et al.’s (2011:499) work on joke-telling, hall writes that “the taking of interactional stances toward [specialized] knowledge may shape distinct identity positions”. as we saw in section 3.2, when interlocutors use the “general kenobi” meme, they not only take the stance that they are “real” star wars fans, they also identify “fake” communal belonging through the stances taken by others. this works largely through the tactics of adequation and distinction (bucholtz & hall 2004), with which members build the identity of a “star wars prequel fan”. for community members, a prequel fan is someone who not only likes the star wars prequel films but can also recognize and participate in recitations of the films’ dialogue. as hall (2019:499) writes, “from an interactional how to do things with memes 9 standpoint, identity emerges within episodes of joke telling as speakers and hearers position themselves in relation to the specialized knowledge they display”. through the specialized knowledge of the scripts and stories of the star wars prequel films, a distinct identity of “star wars prequel fan” emerges, and a community is built through the specific communicative repertoire associated with this identity. this means of building community is well illustrated by the concept of dual indexicality as outlined by jane hill in her work on “mock spanish” (hill 1995). when a member of the star wars prequel meme community makes a joke using a quote from the communal communicative repertoire, they are not only mocking outsiders as “fake”, they are also indexing themselves as a particular kind of humor-loving star wars prequel fan. because the sharing and creating of these prequel memes involves such heavy social implications for community members, they make sure to invoke these references very often, displaying their expertise through their facility with the film scripts and their cleverness in deploying them appropriately. therefore, as seen in the tinder exchanges discussed in section 3.2, many post titles as well as comments on posts consist largely of even more repeated lines from the star wars prequels. here and elsewhere in these online digital communities, when an interlocutor does not understand the reference, they are indexed as “fake” star wars fans. meme creation, as demonstrated throughout this paper, expands upon already circulating communicative repertoires in ways that revitalize the community, modernize its reach, and keep it from “dying out”. for communities built around a specific work of media, memes can additionally foster love and excitement for their chosen work, especially in long-lived series like star wars. although memes have not yet been fully studied in terms of their illocutionary contributions to community-building, the examples analyzed in this paper provide a rich resource for understanding how memetic redistribution, recontextualization, and remediation may serve to create community. references austin, j. l. 1962. how to do things with words. j.o. urmson, & m. sbisà (ed.). cambridge, ma: harvard university press. bauman, richard, & briggs, charles l. 1990. poetics and performance as critical perspectives on language and social life. annual review of anthropology 19.59–88. bolter, jay david, & grusin, richard 1999. remediation: understanding new media. cambridge, ma: mit press. colorado research in linguistics, volume 25 (2021) 10 bucholtz, mary; skapoulli, elena; barnwell, brendan; and lee, jung-eun janie. 2011. entexualized humor in the formation of scientist identities among u.s. undergraduates. anthropology & education quarterly 42(3).177–192. bucholtz, mary, & hall, kira. 2004. language and identity. a. duranti (ed.). a companion to linguistic anthropology. 369-394. hoboken, nj: blackwell. bucholtz, mary, & hall, kira. 2005. identity and interaction: a sociocultural linguistic approach. discourse studies, 7(4-5).585-614. thousand oaks, ca: sage. [hgrant]. 2012. the 18 best chain e-mails you got in 2004. article, 25 january 2012. online: https://www.buzzfeed.com/hgrant/the-18-best-chain-e-mails-you-got-in-2004 hall, kira. 2019. middle class timelines: ethnic humor and sexual modernity in delhi. language in society 48.491–517. hill, jane h. 1995. mock spanish: a site for the indexical reproduction of racism in american english. language & culture. online: https://languageculture.binghamton.edu/symposia/2/part1/index.html [illusions]. 2018. i’m just chillin’ in cedar rapids. youtube video, 11 november 2018. online: https://www.youtube.com/watch?v=nt3vsz2uxj4 iqbal, mansoor. 2021. tinder revenue and usage statistics. business of apps. online: https://www.businessofapps.com/data/tinder-statistics/ edwards, gareth. 2016. rogue one: a star wars story. lucasfilm. kitzinger, celia, and frith, hannah. just say no? the use of conversation analysis in developing a feminist perspective on sexual refusal. discourse & society. 10(3).293–316. lucas, george. 2005. star wars: episode i – the phantom menace. [film]. lucasfilm. lucas, george. 2005. star wars: episode iii – revenge of the sith. [film]. lucasfilm. mcculloch, gretchen. 2019. because internet: understanding the new rules of language. new york: riverhead books. n.a. 2019. drake hotline bling meme generator. online: https://imgflip.com/memetemplate/114388676/inhaling-seagull rymes, betsy. 2012. recontextualizing youtube: from macro-micro to mass-mediated communicative repertoires. anthropology & education quarterly 43(2).214–227. rymes, betsy. 2014. communicating beyond language: everyday encounters with diversity. journal of linguistic anthropology. 24(3)372–374. how to do things with memes 11 [u/ashman508]. 2020. this is where the fun begins! reddit post, 14 february 2020. online: https://www.reddit.com/r/prequelmemes/comments/f3wod6/this_is_where_the_fun_begi ns/ [u/avatyler]. 2017. just how i feel. reddit post, 13 september 2017. online: https://www.reddit.com/r/prequelmemes/comments/6zyac7/just_how_i_feel/ [u/d_feral12]. 2020. i thought she was a star wars fan, guess not. reddit post, 17 january 2020. online: https://www.reddit.com/r/tinder/comments/eqa1jc/i_thought_she_was_a_star_wars_fan_ guess_not/ [u/dhogan73]. 2018. don’t try it. reddit post, 23 january 2020. online: https://www.reddit.com/r/prequelmemes/comments/9bmvai/dont_try_it/ [u/futurarmy]. 2019. oh, it’s just a harmless little bunny, isn’t it boromir?. reddit post, 26 september 2019. online: https://www.reddit.com/r/lotrholygrailmemes/comments/d9ivnn/oh_its_just_a_harmle ss_little_bunny_isnt_it/ [u/notfredrhodes]. 2018. she can’t do that! shoot her…or something! reddit post, 26 june 2018. online: https://www.reddit.com/r/prequelmemes/comments/8u50ud/she_cant_do_that_shoot_her or_something/ [u/poopypants1234321]. 2017. bears, beets, battlestar galactica. reddit post, 25 november 2017. online: https://www.reddit.com/r/prequelmemes/comments/7fidvg/bears_beets_battlestar_galacti ca/ [w__a__c]. 2019. kermit the frog meme template. online: https://imgflip.com/i/3bkinm colorado research in linguistics, volume 25 (2021) 12 endnotes 1 since most digital users on tinder who send messages to women are men, i am using the male pronoun in this section for speakers addressing women. according to iqbal (2021), heterosexual users comprise 88%-99.9% of participants on tinder. microsoft word hodges_cipponeri-cril2021-proof-final.docx 1 how the “law and order” trope individualizes racism and inverts racial vulnerability adam hodges gianna cipponeri university of colorado boulder during the 1968 us presidential campaign, richard nixon infamously ran as the “law and order” candidate, invoking in his republican nomination acceptance speech the domestic protests against racial injustice and the vietnam war. in the 2020 presidential campaign, donald trump revived richard nixon’s “law and order” slogan as part of his response to the black lives matter protests after george floyd’s death in minneapolis on may 25th. in this paper, we examine how trump and his supporters use the “law and order” trope to move public discourse about racism away from critical understandings that view racism as embedded in institutionalized practices and policies, and toward the racial ideology encapsulated in what jane hill (2008) calls the “folk theory of race and racism.” whereas the racial justice movement attempts to center public discourse on systemic racism in policing, the “law and order” trope works to decenter that discourse by individualizing racism and thereby minimizing concerns about the system-wide pattern of racism. as it reinforces the dominant understanding of racism that underpins much us public discourse, the “law and order” trope inverts the racial vulnerability so that black bodies and racial justice protesters are seen as threats rather than victims of state-sanctioned violence. we illustrate these ideas by drawing from examples of public discourse in response to the summer 2020 racial justice protests, including excerpts from tucker carlson and laura ingraham’s shows on fox news in the days immediately following floyd’s killing through ingraham’s interview of trump at the end of the summer. our analysis explains how the discourse spawned by the “law and order” trope reinscribes key assumptions about racism, dismisses calls for racial justice, and perpetuates the racial status quo — thereby posing a substantial barrier to changing the policies and practices that lead to racial inequities in policing. keywords: racism, folk theory of racism, law and order slogan, racial ideology, coded racial appeals 1. the “law and order” trope during the 1968 us presidential campaign, richard nixon infamously ran as the “law and order” candidate, invoking in his republican nomination acceptance speech the domestic protests against racial injustice and the vietnam war (nixon 1968). he talked of “cities enveloped in smoke and flame” and a nation “plagued by unprecedented lawlessness,” juxtaposing those involved in protests with what he called “the forgotten americans — the non-shouters, the nondemonstrators.” nixon deflected accusations that his “law and order” slogan was a “code word for colorado research in linguistics, volume 25 (2021) 2 racism.” those “non-shouters” and “non-demonstrators,” he elaborated, “are not racists or sick; they are not guilty of the crime that plagues the land.” nevertheless, his call “to restore order and respect for law” came to be seen by many as restoring order for white america and respect for laws that continued to unjustly favor white americans at the expense of people of color. as political scientist julia azari remarks, “the question becomes whose order, for whom does the law work” (mcardle 2018). a 1968 cover story in time magazine noted how the phrase was seen as “a shorthand message promising repression of the black community” (waxman 2020). in nixon’s vision of the united states, the law works — and should work — for those non-shouters and non-demonstrators who do not agitate for change, who need not agitate for change because the racial hierarchy works in their favor. but those who do protest the inequities and injustices of a racist system, according to nixon’s logic, are to be considered part of “the criminal forces in this country” (nixon 1968). in the 2020 presidential campaign, donald trump revived nixon’s “law and order” slogan as part of his response to the black lives matter protests after george floyd’s death in minneapolis on may 25th. floyd represents yet another death in a long line of unarmed african americans whose lives have been disproportionately ended by police. the policing of black bodies in public spaces stretches back to the slave patrols of the eighteenth and nineteenth centuries, the terror campaigns and lynchings of the early twentieth century, and continues today even in the more subtle forms of policing of what are implicitly presumed to be white public spaces, giving rise to the colloquial saying, doing x while black, such as driving while black, walking while black, or even birdwatching while black. numerous studies have demonstrated the racial inequities that continue to exist in the criminal justice system (balko 2019), such that black drivers are more likely than white drivers to be stopped and searched (pierson et al. 2020), a significant bias exists in “the killing of unarmed black americans relative to unarmed white americans” (ross 2015), and “young black men are 21 times as likely as their white peers to be killed by police” (gabrielson, jones, and sagara 2014). to ignore the presence of racism in today’s society requires suppressing the historical throughline from the slave patrols to the continued criminalization of black bodies in public spaces that results in the deaths of those like george floyd. it requires a willful ignorance of those histories and current realities that shape racism in contemporary us society (mills 2008). but how how the “law and order” trope individualizes racism and inverts racial vulnerability 3 is this discursively achieved and how is the “law and order” trope leveraged to dismiss calls for racial justice, as was done in the summer of 2020? we argue that the “law and order” trope operates by moving the focus away from critical understandings of racism as embedded in institutionalized practices and policies, and toward the racial ideology encapsulated in what jane hill (2008) calls the “folk theory of race and racism.” whereas the racial justice movement attempts to center public discourse on systemic racism in policing, the “law and order” trope works to decenter that discourse by individualizing racism and thereby minimizing concerns about the system-wide pattern of racism. as it reinforces the dominant understanding of racism that underpins much us public discourse, the “law and order” trope inverts the racial vulnerability so that black bodies and racial justice protesters are seen as threats rather than victims of state-sanctioned violence. we illustrate these ideas by drawing from examples of public discourse in response to the summer 2020 racial justice protests.1 2. the coded racial appeals of the “law and order” trope on the surface, overt appeals for “law and order” are couched as benign calls for social order; but the slippage between the dual senses of social order paves the way for a defense of the racial status quo. the first sense of the term social order refers to orderliness in contrast to unrest (social order1). in much of the “law and order” discourse, appeals to “law and order” are juxtaposed with images of street protests marked by descriptors such as “chaos” and “unrest.” for example, on may 29, days after george floyd was killed, fox news host tucker carlson opened his show with a focus on the protests, saying, “remarkable scenes of violence and destruction and chaos from across the country now and we’re going to spend much of the hour keeping you abreast of what’s happening” (see appendix a for the full excerpt). in the beginning of her interview with donald trump at the end of the summer, fox news host laura ingraham prefaced a question to trump by saying, “so when you see the unrest on the streets — and so much of it is driven by an antipathy toward law enforcement” (see appendix d for the full excerpt). in both examples, the descriptors “chaos” and “unrest” paint the protests for racial justice as the antithesis of social order in the first sense of the term. this allows the protests to be loosely glossed under the rubric of sowing disorder and lawlessness. colorado research in linguistics, volume 25 (2021) 4 the overwhelming emphasis placed on disorder in the streets necessitates a response that would restore orderliness (social order1). it does this by couching the appeal for “law and order” within a commonly accepted understanding and desire for social order — as opposed to social unrest. the discourse plays up incidents of violence that accompany the protests — positioning even the peaceful protests as inherently disorderly — and downplaying (or simply ignoring) the complaints about systemic racism that underpin the protests. but while the surface appeal to social order (social order1 in contrast to social unrest) may fall within the general moral order, making it easy to accept for uncritical listeners, the second sense of social order refers to a system of social structures, institutions, and practices. the us system is based on a racial order that organizes the differential distribution of justice according to the racial hierarchy. that racial order (social order2) represents the racial status quo that has become so problematic for many americans, leading people into the streets to protest the injustices it spawns. the coded racial appeal of the “law and order” trope arises from this slippage between the first sense of social order (as orderliness in contrast to unrest) and the second sense (as the system that represents the current racial order). discussants can use the “law and order” trope to ostensibly talk about countering social unrest while also implicitly defending the current racial order. this coded message enables people to take a pro “law and order” stance under the guise of supporting social orderliness while covertly signaling their support for the racial status quo. although the racial appeals are mostly covert, the dual senses of social order are frequently invoked so that the “law and order” discourse often becomes about more than simply restoring orderliness (social order1); it is also about protecting the system (social order2). for example, in the opening monologue to his may 29 show, tucker carlson declares, “what you're watching is the ancient battle between those who have a stake in society and would like to preserve it, and those who don't and seek to destroy it” (see appendix a for the full excerpt). in these remarks, carlson suggests that the threat involves not just disruptions to orderliness in the streets (social order1), but that society itself (social order2) is being threatened. in her show on june 1, a week after george floyd’s killing, laura ingraham likewise reframes the outrage over the injustice inflicted on george floyd to position the protesters as wanting to, in her words, “murder america.” she says, “all people of good faith agree that what happened to george floyd was heinous and depraved. it was murder. but that's not what we're seeing on our violent streets. we're not seeing outrage really expressed about that. and that's not what the how the “law and order” trope individualizes racism and inverts racial vulnerability 5 criminals and the domestic terrorists are perpetrating as they use mr. floyd's killing to try to murder america” (see appendix c for the full excerpt). as seen in these examples, the discourse associated with the “law and order” trope works to remove the motive of the protesters so that any societal-wide racial justice advocacy is seen as part and parcel of a movement of those who, in ingraham’s words, “try to murder america,” or in carlson’s words, “seek to destroy” society (social order2). by arguing that the motivation behind the demonstrations is disingenuous, they establish grounds to identify racial justice protesters as criminals. the appeal to “law and order” thereby becomes a defense of the racial order through the delegitimization of the racial justice protesters and their concerns, allowing those wielding the “law and order” slogan to dismiss those concerns without overtly negating the protest mantra that “black lives matter.” insofar as protesters and racial justice advocates want to change a racially unjust system, carlson and ingraham are right to see that system (social order2) as being challenged. but their presentation of the situation fails to acknowledge the reason for that challenge (systemic racism) and instead moves to an all-out defense of the system (racism and all). as discussed in the next section, this failure for those operating within the “law and order” discourse to recognize the protesters’ concerns stems from the dominant racial ideology that individualizes racism and ignores it as a systemic problem. 3. the racial ideology of the “law and order” discourse anthropologists widely recognize that race is a cultural construct. as audrey smedley (2007) explains, “race originated as a folk idea and ideology about human differences; it was a social invention, not a product of science” (2). but by the end of the eighteenth century, those folk ideas began to be propped up by scientific and pseudo-scientific techniques that “sought to affirm the differences between blacks and whites” (smedley 2007: 7). this cultural project helped justify and rationalize the enslavement of those racialized as black within the us context. folk ideas continue to underpin popular understandings of race and racism. jane hill (2008) encapsulates these ideas in what she terms the “folk theory of race and racism.” as the dominant racial ideology in us society, the folk theory provides “the racially based frameworks” that explain, justify, and defend “the racial status quo” (bonilla-silva 2006: 9). central to this dominant ideology is an incomplete recognition of racism, locating it merely in “individual beliefs, colorado research in linguistics, volume 25 (2021) 6 intentions, and actions” (hill 2008: 6) and erasing how it operates as a system of power to structure the social hierarchy and differentially distribute justice according to that hierarchy. drawing from this ideology, the “law and order” discourse works to dismiss the racial justice movement’s focus on systemic racism in policing by individualizing incidents like the killing of george floyd — positioning such incidents as one-off events perpetrated by individual outliers within a system otherwise untainted by racist policies and practices. the individualization of such events is accomplished through the ideological process of erasure. as judith irvine and sue gal (2000) explain, “erasure is the process in which ideology...renders some persons or activities...invisible. facts that are inconsistent with the ideological scheme either go unnoticed or get explained away” (38). erasure can be seen in practice in the interview fox news host laura ingraham conducted with president trump at the end of the summer of 2020 (see appendix d for the full excerpt). in the interview, ingraham asks trump about “the statistics that are cited over and over again” of more african americans being stopped by police. trump invokes the “law and order” trope in his response, stating, “what the black community wants in this country is they want police and they want law and order. [...] look, they want law and order. they want the police.” he goes on to say, “they've gotten along with the police, and the police have been very badly mistreated because you got one bad apple, and it becomes a story for weeks.” the bad apple metaphor invoked by trump is commonly used to individualize racist acts of police violence, positioning officers involved in the killing of unarmed african americans as outliers. this perspective accords with the folk theory’s ideology, which identifies “racists” as individual outliers who engage in isolated acts of bigotry. this distances the individual actions from society writ large, rendering invisible the way those individual actions are part of a broader pattern of discriminatory actions and policies that disproportionately impact african americans. as the interview continues, trump reiterates the bad apple metaphor as he compares the bad apple to a golfer who simply makes a mistake and misses a shot. he says, “the police are under siege because of things — they can do 10,000 great acts, which is what they do, and one bad apple, or a choker — you know a choker, they choke — shooting the guy in the back many times.” interestingly, trump’s reference to “one bad apple” remains unspecified in the conversation. he may be referring to the officer who killed george floyd or, more likely in the last reference about “shooting the guy in the back many times,” he probably has in mind the atlanta officer who how the “law and order” trope individualizes racism and inverts racial vulnerability 7 shot and killed rayshard brooks a few weeks after floyd was killed. the fact that the “one bad apple” could refer to any number of officer shootings of african americans within the months prior to his interview — for example, george floyd, rayshard brooks, breona taylor, daniel prude, jacob blake — underscores the very pattern that the individualization of those acts works to erase. speaking to fox news host tucker carlson a few days after george floyd’s death, sen. ted cruz starts by acknowledging the injustice (see appendix b for the full excerpt): “well, listen, it's horrific and it starts with a horrific act of police brutality.” a few moments later, he underscores that the incident was carried out by a single individual as he objects to those trying to focus on the larger societal pattern; they “want to use this incident of clear abuse by one police officer and they want to use it to paint every police officer as corrupt and racist,” cruz says. according to the ideology that supports the “law and order” discourse, cruz can recognize the injustice of a single incident of police brutality (as did ingraham in an earlier excerpt), but in doing so he needs to disconnect that incident from the larger pattern. the erasure of the system-wide pattern of statesanctioned violence against african americans is part of what eduardo bonilla-silva (2013) refers to as the minimization of racism, regarding “discrimination exclusively as all-out racist behavior” while insisting that a societal pattern of “discrimination is no longer a central factor affecting minorities’ life chances” (29). this allows observers to recognize the injustice inflicted upon george floyd while still negating that the incident fits a wider pattern. 4. the outcomes of the “law and order” discourse as the calls for racial justice during the summer of 2020 shined a spotlight on systemic patterns of state-sanctioned violence against african americans, much of the public discourse centered on trying to reckon with the societal problem of racial disparities within policing and the criminal justice system. that reckoning, however, posed a substantial challenge to those committed to the racial status quo. the response, animated by trump and many of his allies, was to recycle nixon’s “law and order” slogan. calls for “law and order” ostensibly sound as if they fall within the general moral order, somewhat similar to the idea that “black lives matter.” but beneath the vagueness of this simple slogan resides the pernicious logic of the dominant racial ideology. drawing from the folk ideology, the “law and order” discourse isolates incidents like the killing of george floyd as outliers having nothing to do with policing in general. this minimizes the colorado research in linguistics, volume 25 (2021) 8 system-wide problem of policies and practices that contribute to a pattern of incidents of which george floyd is but one example. discounting the problem of systemic racism in policing also works to remove the motive of the racial justice protesters. if, according to the “law and order” logic, there is no systemic racism; then the actions of the racial justice movement to focus attention on systemic issues have little or nothing to do with justice for george floyd. as tucker carlson says a few days after floyd’s death while pointing to incidents of burning and looting, “underneath it all, this violence doesn't have that much to do with the behavior of the minneapolis police department” (appendix a). the frequent representation of racial justice protesters through such images paints all street protests in a similar light of criminality. the criminalization of the protesters inverts the racial vulnerability felt by african americans at the hands of the police by positioning the police as the victims of racial justice protesters and the movement to affirm that black lives matter. this is illustrated in the ingraham-trump interview in which the two speakers co-construct a narrative where, in trump’s words, “police are under siege” (appendix d). the police are the ones said to be under threat, rather than recognizing the threat to black lives and bodies within a system that continues to operate as if black lives do not matter. the “law and order” slogan effectively counters the critical focus placed on racism in policing by inverting the threat and absolving the institution of policing and the wider social order as having nothing to do with each new incident of police brutality. the discourse spawned by the “law and order” trope thereby reinscribes key assumptions about racism, dismisses calls for racial justice, and perpetuates the racial status quo. the discourse, especially as it subtly works to defend the racial hierarchy, represents a substantial barrier to the types of systemic change needed to overcome the racist policies and practices entrenched in the system. references balko, radley. 2019, april 9. “21 more studies showing racial disparities in the criminal justice system.” the washington post. https://www.washingtonpost.com/opinions/2019/04/09/more-studies-showing-racialdisparities-criminal-justice-system/ bonilla-silva, eduardo. 2006. racism without racists. rowman & littlefield. gabrielson, ryan, ryann grochowski jones, and eric sagara. 2014, october 10. “deadly force, in black and white.” propublica. https://www.propublica.org/article/deadly-force-inblack-and-white. how the “law and order” trope individualizes racism and inverts racial vulnerability 9 gal, susan and judith t. irvine. 2000. “language ideology and linguistic differentiation.” in p. kroskrity’s regimes of language: ideologies, polities, and identities, pgs. 35-83. santa fe, nm: school of american research. hill, jane. 2008. the everyday language of white racism. malden, ma: wiley-blackwell. mcardle, terence. 2018, november 5. “the 'law and order’ campaign that won richard nixon the white house 50 years ago.” the washington post. https://www.washingtonpost.com/history/2018/11/05/law-order-campaign-that-wonrichard-nixon-white-house-years-ago/ mills, charles. 2008. “white ignorance.” in agnotology: the making and unmaking of ignorance, robert proctor and londa l. schiebinger (eds.), 230-249. stanford: stanford university press. nixon, richard. 1968, august 8. “address accepting the presidential nomination at the republican national convention in miami beach, florida.” the american presidency project, uc santa barbara. https://www.presidency.ucsb.edu/documents/addressaccepting-the-presidential-nomination-the-republican-national-convention-miami pierson, emma; camelia simoiu; jan overgoor; sam corbett-davies; daniel jenson; amy shoemaker ; vignesh ramachandran; phoebe barghouty; cheryl phillips; ravi shroff; and sharad goe. 2020. “a large-scale analysis of racial disparities in police stops across the united states.” nature human behaviour 4: 736-745. https://doi.org/10.1038/s41562020-0858-1 ross, cody t. 2015, november 5. "a multi-level bayesian analysis of racial bias in police shootings at the county-level in the united states, 2011–2014." plos one. https://doi.org/10.1371/journal.pone.0141854 smedley, audrey. 2007, march 14-17. “the history of the idea of race...and why it matters.” paper presented at the conference, “race, human variation and disease: consensus and frontiers,” sponsored by the american anthropological association. warrenton, va. https://understandingrace.org/resources/pdf/disease/smedley.pdf waxman, olivia b. 2020, june 2. “trump declared himself the 'president of law and order.' here's what people get wrong about the origins of that idea.” time. https://time.com/5846321/nixon-trump-law-and-order-history/ colorado research in linguistics, volume 25 (2021) 10 endnotes 1 discourse excerpts come from cable news transcripts provided by lexisnexis. we used the search term “law and order” to search the lexisnexis database for transcripts from fox news and cnn between the dates april 1, 2020 and august 31, 2020. additional context for excerpts quoted in the paper can be found in the appendices where full transcript excerpts are provided. how the “law and order” trope individualizes racism and inverts racial vulnerability 11 appendix a “tucker carlson tonight” (fox news) may 29, 2020 carlson: mike tobin for us in minneapolis. thank you. remarkable scenes of violence and destruction and chaos from across the country now and we're going to spend much of the hour keeping you abreast of what's happening. but before we dive into that, we want to focus on one single thing that happened last night. a police station in a major american city was occupied and looted and burned. most of us assumed we would never live to see something like that happen here, but it did happen. so, the question is, has anyone been arrested for doing that? will anyone ever be arrested? no one in authority seems especially interested in apprehending the people who did it. all of it happened on camera, but the perpetrators just walked away, and it's possible maybe likely that most of them will never be punished for it. that's striking. it's a very different experience from the ones most americans have living here. as minneapolis burns and crowds grow in the streets of atlanta and many other cities, the rest of us are continuing on as we always do. dutifully following the rules. there are many of those. every year, there seems to be countless new rules to follow. they multiply like insects. we do our best to keep up. we get our permits, apply for our licenses, put on our reading glasses and check the latest regulations on the internet. we wear our little masks. we keep our dogs on leashes. we drive sober. we don't eat on the subway. we never litter. colorado research in linguistics, volume 25 (2021) 12 we make orderly lines and patiently wait our turn. in airports and government buildings, we remove our shoes and submit to body searches from strangers. we lose our dignity every time we do this, but they tell us we must, so we accept it without complaint. in public, we hide what we really think. we bury our natural instincts. we keep our deepest beliefs to ourselves. we know the boundaries. we understand we will be punished for telling the truth. this is the america the rest of us live in. for the privilege of citizenship in a country like this, we work as hard as we can. we never stopped sharing what we earn with others. we send money we would rather give to our own children to politicians in faraway cities. with that money, they make new rules. we follow those rules to the letter. that's what we were told to do as children. that's the deal we've struck, at least, we thought it was. now, we know that other people have somehow negotiated a far better deal than the one we have. they get to ignore the rules. they don't believe in order or fairness. they reject society itself. reason and process and precedent mean nothing to them. they use violence to get what they want, immediately. people like this don't bother to work. they don't volunteer or pay taxes to help other people. they live for themselves. they do exactly what they feel like doing. they say exactly what they feel like saying. they spray paint their opinions on buildings. how the “law and order” trope individualizes racism and inverts racial vulnerability 13 on television, hour by hour, watch these people, criminal mobs destroy what the rest of us have built. they have no right to do that. they don't contribute to the common good, they never have. yet, suddenly they seem to have all the power. this is hardly the first time something like this has happened in america. spasms of destructive violence, a recurring feature of our history, in fact of every country's history. the ideologues will tell you that the problem is race relations or capitalism or police brutality or global warming, but only on the surface. the real cause is deeper than that, and it's far darker. what you're watching is the ancient battle between those who have a stake in society and would like to preserve it, and those who don't and seek to destroy it. underneath it all, this violence doesn't have that much to do with the behavior of the minneapolis police department. for evidence, watch this tape. it's from the 1992 riots in los angeles. it was shot almost 30 years ago. it could have been shot this afternoon. colorado research in linguistics, volume 25 (2021) 14 appendix b “tucker carlson tonight” (fox news) may 29, 2020 carlson: so i'm going to ask you the question that i asked deroy, you're watching these pictures. you followed this for the past three days. what do you make of this? where do you think it is going? cruz: well, listen, it's horrific and it starts with a horrific act of police brutality. you know, anytime you have a police officer involved shooting, the media often goes into a frenzy and there is an immediate demonization and attack of the police officers, and i think that's wrong. i think it's premature. that being said, in this instance, we have a video of the incident and we can see with mr. floyd, the officer with his knee on his neck for eight minutes. mr. floyd has his handcuffs. he is clearly incapacitated. he is begging for his life -and what we saw was wrong. there's no legitimate law enforcement purpose for what we saw right there. carlson: well, i'm sorry, senator, let me just -let me just stop right there and just ask a question, and i agree with you, i found the video very upsetting. i mean, there's a lot of abuse of power by a lot of different people in charge, including the police sometimes. do you believe, since let's just deal with facts here -do you believe that the man in custody died of suffocation because the police officer was sitting on him? cruz: i don't know. we'll have to see what the medical evidence shows. but what i do - carlson: wait. wait. hold on. isn't that the question? i mean, either the cop killed him or he didn't? i mean -no? cruz: the question is, was that abuse of authority and police brutality? carlson: it was definitely an abuse of authority. but the guy has been charged with murder. so, isn't the question whether he killed him or not? how the “law and order” trope individualizes racism and inverts racial vulnerability 15 cruz: well, there are two separate questions. number one, the department of justice opened a civil rights investigation. that was the right thing to do. i applauded the department of justice for doing that. number two, the prosecutor chose today to bring homicide charges. now, to prove that, they will have to prove that the evidence supports it. i don't know what the medical examiner is going to determine on that. so, whether or not it was homicide will depend on the evidence, but it was clearly police brutality, and it was not conduct we expect of any officer. the officers are entitled to defend themselves - carlson: i totally agree with that. i think it was awful. i've seen that kind of -i covered cops. i've seen the kind of thing before and i hate it. however, the country is convulsing on the basis of the idea that a cop killed this man who was restrained. he was in handcuffs. and i just -i think it's a meaningful question. it's not something we could alight over them and like, oh, it doesn't matter. of course it matters. why wouldn't it matter? cruz: tucker, saying that the criminal justice system will operate and it will depend upon what the evidence is and whether the case be proven to the jury is not alighting over it. it's saying - carlson: no, it's not and i agree with you a hundred percent there. cruz: and, and one of the reasons, sadly, that we are seeing this violence and this rioting is that you have a lot of demagogues that want to use this incident of clear abuse by one police officer and they want to use it to paint every police officer as corrupt and racist. and most police officers heroically risked their lives to protect the communities they're in, often minority communities and for everyone that is stirring up racial division and engaging in violence and looting, that is completely unacceptable. violence and criminal conduct is unacceptable whether it is committed by a mob in rage or whether it's committed by a police officer who is breaking the law. the law should apply fairly and uniformly to everyone. colorado research in linguistics, volume 25 (2021) 16 carlson: so, why is that so difficult for so many republicans in washington to say, i saw the tape, i was horrified by it. the guy should be punished for doing this, and you're not allowed to burn our cities down. how the “law and order” trope individualizes racism and inverts racial vulnerability 17 appendix c “the ingraham angle” (fox news) june 1, 2020 ingraham: we're going to check back with you as 11 pm draws closer. now, as i said we're going to get to the action on some streets across the nation in moments. but first, this is my statement about what happened today and what's been happening. restoring order. that's the focus of tonight's angle. all people of good faith agree that what happened to george floyd was heinous and depraved. it was murder. but that's not what we're seeing on our violent streets. we're not seeing outrage really expressed about that. and that's not what the criminals and the domestic terrorists are perpetrating as they use mr. floyd's killing to try to murder america. appendix d “the ingraham angle” (fox news) august 31, 2020 ingraham: so when you see the unrest on the streets — and so much of it is driven by an antipathy toward law enforcement. trump: yes. ingraham: and more african-americans are stopped by the police, the statistics that are cited over and over again. what can you say to those families who live on those streets and who are worried? they're worried because they think their sons or even- trump: yes. ingraham: --their daughters could be targeted. because i know because i've known you for a long time, you don't want that. you want people to all be treated equally. but they have a caricature of republican voters, and you're the leader of the party. what do you say to them about that mischaracterization? colorado research in linguistics, volume 25 (2021) 18 trump: what the black community wants in this country is they want police and they want law and order. they don't want what's happening to their communities. they're being affected in a much harsher, meaner manner than anybody else. that includes hispanics, where i'm doing very well also. look, they want law and order. they want the police. they do polls, and the polls are at 82, 83 percent, they want the police. they've gotten along with the police, and the police have been very badly mistreated because you got one bad apple, and it becomes a story for weeks. ingraham: st. louis african-american police officer shot in the head and killed - trump: yes, dorn. ingraham: -last night. no another african-american just killed yesterday. trump: that's true. yes, that's true. killed. ingraham: it's more dangerous to be a police officer today, do you not think, than it has been a long time? trump: the police are under siege because of things -they can do 10,000 great acts, which is what they do, and one bad apple, or a choker -you know a choker, they choke -shooting the guy in the back many times. couldn't you have done something different? couldn't you have wrestled him? in the meantime, he might've been going for a weapon. and there's a whole big thing there. but they choke. just like in a golf tournament, they miss a three-foot - ingraham: you're not comparing it to golf, because of course that's what the media - trump: i'm saying people choke. ingraham: people make -people panic. how the “law and order” trope individualizes racism and inverts racial vulnerability 19 trump: people choke. and people are bad people. you have both. you have some bad people, and they choke. you could be a police officer for 15 years, and all of a sudden you're confronted. you've got a quarter of a second to make a decision. if you don't make the decision and you're wrong, you're dead. people choke under those circumstances, and they make a bad decision. i've seen bad decisions of people that it looked bad but probably it was a choke. but you also have bad police, but you also, the vast -not only the vast majority -thousands and thousands of great acts, and one bad one, and you make the evening news for weeks. the dialogic emergence of ‘truth’ in politics: reproduction and subversion of the ‘war on terror’ discourse colorado research in linguistics. june 2008. vol. 21. boulder: university of colorado. © 2008 by adam hodges. the dialogic emergence of ‘truth’ in politics: reproduction and subversion of the ‘war on terror’ discourse adam hodges university of colorado truth claims in political discourse are implicated in a dialogic process whereby political actors "assimilate, rework, and re-accentuate" prior discourse (bakhtin 1986:89). while political actors themselves may view truth as an object to be discovered, i argue that discourse analysts are best served by viewing truth as an emergent property of this dialogic process. in this paper, i examine how intertextual connections are integral to both the reproduction and subversion of established truth claims (such as the claim that saddam hussein possessed weapons of mass destruction). my data draw from george w. bush's speech on may 1, 2003 to declare the end of "major combat operations" in iraq, the first presidential debate between john f. kerry and george w. bush in september 2004, and joseph lowery's speech during the coretta scott king funeral in february 2006. my analysis examines these data in light of key phrases (e.g. "weapons of mass destruction") that form intertextual series across these contexts, as well as the role of reported speech in connecting one discursive encounter with another. 1. introduction in american political discourse, debates are often framed around issues of truth. in recent years, questions surrounding the possession of weapons of mass destruction by saddam hussein or the culpability of iraq in the events of 9/11 have taken center stage. political actors wield facts to show that they have uncovered what they claim to be the ‘real’ truth as they counter their opponents’ truth claims. yet, as widely recognized by postmodern philosophers, truth in these debates is not so much discovered as enacted. that is, truth is not simply an object external to the debate; but rather, a form of knowledge emergent from the debate.1 a version of this paper was presented at the 2007 culture, language, and social practice conference at cu-boulder. i would like to thank the organizers and participants for the invaluable discussions that resulted from the conference. 1 not surprisingly, opposing sides in debates both feel they have ‘truth’ on their side. when our side believes we have the truth, but our opponents do not, we usually say one of two things: either our opponents simply lack appropriate information—that is, they haven’t seen all the facts yet. or to be less generous, we might accuse them of knowing the truth (like we do) but of obscuring it because it is damaging to their cause—that is, they must be lying (cf. jervis 2006). in either case, political actors still orient to truth as an object. truth is an object wielded in political 1 hodges: the dialogic emergence of ‘truth’ in politics published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 2 while political actors themselves may view truth as an object to be discovered, i argue that discourse analysts are best served by viewing truth as an emergent property of a dialogic process. in describing austin’s (1962) work on performative utterances, duranti (1993) notes that “words do not simply describe the world, they also change it” (235). to take this further, words do not simply describe a pre-existing truth; words in political discourse effectively help realize it. so as analysts, we need to place our focus on “the emergent, interactive nature of the process” (duranti 1993: 227). in this paper, i examine political discourse as a dialogic process through a framework that draws upon the bakhtinian-inspired idea of intertextuality (bakhtin 1981, 1986; kristeva 1980). social actors do not formulate utterances in a vacuum; nor do individual “speech events” (jakobson 1960; hymes 1974) take place in isolation from one another. rather, as bauman and briggs (1990) note, “a given performance is tied to a number of speech events that precede and succeed it” (60). discourse, according to bakhtin, “cannot fail to be oriented toward the ‘already uttered,’ the ‘already known,’ the ‘common opinion’ and so forth” (bakhtin 1981: 279).2 2. repetition and variation on a theme the concept of intertextuality is important in understanding the dialogic emergence of truth because it allows the analyst to do more than describe the structure of discourse in isolation, and instead to connect it with the larger interpretive web in which it is embedded. the interconnectivity of discourse is central to both the reproduction of truth claims as well as the subversion of truth claims. for truth claims to become widely accepted as valid and credible versions of reality, they must enter into the public domain where they are repeated, reaffirmed, and reified. even in the challenging of established truth claims, political actors do not create utterances completely from scratch, but rather construct their utterances out of a reservoir of prior discourse. therefore, political actors involved in either the reproduction or subversion of truth claims draw from previously uttered words, which, as bakhtin (1986) describes, they “assimilate, rework, and re-accentuate” (89). in the examples that follow, i focus on the way key phrases are reiterated across different types of contexts by both george w. bush and his political opponents. as kristeva (1980) points out, the repetition of prior text may be done “seriously, claiming and appropriating it without relativizing it” or the process of debate, but it is an object that is seen as separate from (outside of) political debate itself. (see also duranti 1993 for a discussion of truth and intentionality). 2 as mannheim and tedlock (1995) state, “any and all present discourse is already replete with echoes, allusions, paraphrases, and outright quotations of prior discourse” (7). 2 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/1 doi: https://doi.org/10.25810/gkzf-q768 the dialogic emergence of ‘truth’ in politics 3 recontextualization may introduce “a signification opposed to that of the other’s word” (73). in its extreme, resignification may move into the realm of parody (bakhtin 1981: 340; cf. álvarez-cácamo 1996: 38). i explore each of these dimensions in turn. 3. establishing and reinforcing an intertextual series the following two examples illustrate the use of talking points to reinforce the bush ‘war on terror’ narrative, a narrative that forwards a powerful set of assumptions and explanations about america’s struggle against terrorism since 9/11 (hodges 2007). central to this narrative is the truth claim surrounding the presence of weapons of mass destruction (wmds) in iraq. saddam hussein’s supposed possession of wmds is made doubly worrisome in the narrative due to a second truth claim about his putative ties to terrorist organizations. excerpts (1) and (2) are both taken from george w. bush’s speech on may 1, 2003 aboard an aircraft carrier off the coast of san diego. in that speech, he declared that “major combat operations in iraq have ended.” in these examples, we see the reiteration of talking points that reinforce the truth claims in the ‘war on terror’ narrative. in particular, we see the key phrase “weapons of mass destruction” embedded in this discourse (underlined in the examples). 1) from bush’s speech on the end of major combat operations in iraq, may 1, 2003 the liberation of iraq is a crucial advance in the campaign against terror. we've removed an ally of al qaeda, and cut off a source of terrorist funding. and this much is certain: no terrorist network will gain weapons of mass destruction from the iraqi regime, because the regime is no more. ((applause)) 2) from bush’s speech on the end of major combat operations in iraq, may 1, 2003 any outlaw regime that has ties to terrorist groups and seeks or possesses weapons of mass destruction is a grave danger to the civilized world -and will be confronted. ((applause)) the phrase “weapons of mass destruction” forms part of an intertextual series (hanks 1986; cf. hill 2005). as it enters into subsequent contexts, it points back to the prior contexts where it has been previously uttered. namely, this includes numerous presidential speeches prior to and after the invasion of iraq where the phrase is embedded in truth claims about the threat posed by saddam hussein. moreover, the diachronic repetition (tannen 1989)3 of this phrase occurs in sound bites from these speeches that are recontextualized in media reportage that reiterates these truth claims. in this way, an important indexical association is 3 tannen (1989) use the term ‘diachronic repetition’ to refer to intertextuality, as opposed to ‘synchronic repetition,’ or intratextuality (i.e. repetition within a text). 3 hodges: the dialogic emergence of ‘truth’ in politics published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 4 formed between this key phrase and the contexts where the ‘war on terror’ narrative is articulated. as developed by charles peirce and further refined by ochs (1992, inter alia), silverstein (1976, 1985, inter alia) and others, the notion of indexicality “is the semiotic operation of juxtaposition” (bucholtz and hall 2004: 378) whereby contiguity is established between a sign and its meaning. as bauman reminds us, indexicality is important in intertextual connections. he notes that: bakhtin’s abiding concern was with dimensions and dynamics of speech indexicality— ways that the now-said reaches back to and somehow incorporates or resonates with the already-said and reaches ahead to, anticipates, and somehow incorporates the to-be-said. (bauman 2005: 145) in sum, the repeated juxtaposition of the phrase “weapons of mass destruction” in contexts where bush reiterates the truth claims in the ‘war on terror’ narrative allows this phrase to effectively operate as an index for those claims. in excerpt (3), taken from the first presidential debate between bush and john f. kerry before the 2004 election, we again see the repetition of this key phrase as bush reiterates elements of his narrative. 3) from the first presidential debate, september 30, 2004 bush: we're facing a group of folks who have such hatred in their heart they'll strike anywhere, with any means. and that's why it's essential that we have strong alliances, and we do. that's why it's essential that we make sure that we keep weapons of mass destruction out of the hands of people like al qaeda, which we are. in these first three examples, we see repetition of an intertext done in a manner that, as kristeva (1980: 73) points out, takes what is repeated seriously, without relativizing it. in fact, the recontextualization of the phrase “weapons of mass destruction,” lifted by bush out of prior presidential speeches and placed into subsequent contexts such as the debate, merely works to reinforce its previously established social meaning. in other words, the phrase indexes and attempts to bolster the truth claims espoused by the administration. even in contexts such as the debate where the entire bush ‘war on terror’ narrative may not be told in full detail, the invocation of the key phrase may be sufficient to point to it and thereby work to reinforce, or at least remind an audience of its claims. 4. recontextualization and the reshaping of prior text the process of lifting key phrases out of one context and moving them to another allows social actors to bring with the text varying degrees of the earlier context while also transforming the text in the new setting (cf. gal 2006: 178; 4 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/1 doi: https://doi.org/10.25810/gkzf-q768 the dialogic emergence of ‘truth’ in politics 5 voloshinov 1971). while the indexical associations between a key phrase and its contextual significance may draw on already established meanings—what silverstein (2003) terms presupposed indexicality—new indexical links may also be created—what silverstein terms creative or entailed indexicality.4 put another way, the social meanings associated with an indexical sign are both partly preestablished and partly recalibrated when that sign is brought into a new context.5 in this way, prior text is always open to reinterpretation and reshaping as it enters into new settings. political discourse, in particular, is effectively a struggle over entextualization. it is a struggle over whose “preferred reading” of a prior text will be accepted as more valid (cf. blommaert 2005: 47). control over the process of entextualization is frequently achieved through the use of reported speech. voloshinov (1973) provides a significant discussion on the topic where he characterizes reported speech as “speech within speech, utterance within utterance, and at the same time also speech about speech, utterance about utterance” (115; italics in original). voloshinov’s comments highlight the capacity of reported speech to not just represent pieces of previously uttered discourse, but to re-present what has been said elsewhere by others—that is, to effectively recontextualize a prior utterance with “varying degrees of reinterpretation” (bakhtin 1986: 91). voloshinov (1973) explains that the use of reported speech “imposes upon the reported utterance its own accents, which collide and interfere with the accents in the reported utterance” (154). as buttny (1997) summarizes, “reporting speech is not a neutral, disinterested activity. persons report speech along with assessing or evaluating it” (484). 5. reported speech in the challenging of truth claims the next example also comes from the first 2004 presidential debate. in (4), kerry uses a reported speech frame to attribute and re-present words previously uttered by bush. (the quotatives are highlighted in bold and the reported words are underlined.) 4 as silverstein (2003) explains, “any socially conventional indexical” sign is “dialectically balanced between” what he calls indexical presupposition and indexical entailment (195). 5 we might think of this recalibration in terms of social meanings as emergent properties of interaction: the meanings that emerge may simply reaffirm established ones or may involve significant modifications made within the current context. social meanings are never fixed once and for all, but are subject to continual renewal through micro-level discursive encounters; and therein exists the potential for shifts in indexical associations, which in turn contribute to changes in macro-level social categories and forms of knowledge. silverstein (2003) stresses that indexicality should not be misconstrued “as being micro-contextually deterministic” (197). 5 hodges: the dialogic emergence of ‘truth’ in politics published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 6 4) from the first presidential debate, september 30, 2004 kerry: now, i'd like to come back for a quick moment if i can to that issue about china and the talks because that's one of the most critical issues here -north korea. just because the president says it can't be done, that you'd lose china, doesn't mean it can't be done. i meanthis is the president who said there were weapons of mass destruction, said mission accomplished, said we could fight the war on the cheap, none of which were true. we can have bilateral talks with kim jong il and we can get those weapons at the same time as we get china because china has an interest in the outcome too. in the highlighted portion of this example, kerry begins by reanimating the phrase “weapons of mass destruction” that we have already seen in bush’s discourse. by bringing this phrase into the context of the debate, kerry invokes the truth claims forwarded by bush. yet kerry does not bring these words into the new setting to simply maintain fidelity to the way these words have been used previously by bush. rather, the reported speech frame allows kerry to reshape the words in line with a different interpretation. in his metapragmatic comments about these reported words (italicized in the example), kerry provides his own interpretation about their larger significance in the debate over iraq and terrorism. as sacks (1992) points out, the reported speech frame works to convey to listeners “how to read what they’re being told” (274; cited in buttny 1998: 49). in other words, reported speech frames can recontextualize another’s words in line with the present speaker’s desired interpretations. importantly, this reshaping of prior text works to recalibrate the larger social meanings associated with the phrase “weapons of mass destruction.” instead of simply indexing the truth claims espoused by bush, the phrase now begins to form an association with an alternative narrative which undermines the veracity of those claims and links the phrase “weapons of mass destruction” with a deceptive policy put forth by the administration. next, kerry reiterates another key phrase from bush’s prior discourse: “mission accomplished.” this phrase stands metonymically for the event aboard the aircraft carrier on may 1, 2003 where bush declared the end of major combat operations in iraq. (recall that the first two examples were drawn from this speech.) the phrase “mission accomplished” was prominently displayed on a banner behind the podium where bush spoke, pictured in (5). 6 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/1 doi: https://doi.org/10.25810/gkzf-q768 the dialogic emergence of ‘truth’ in politics 7 5) bush’s declaration of the end of major combat operations in iraq, may 1, 2003 one might compare these economical references to keith basso’s description of the way apaches speak with place-names. as basso (1996) describes, apaches often invoke a particular place-name in the midst of conversation to conjure up a shared narrative associated with that place. without reiterating the narrative itself, mentioning the place-name is sufficient to set interlocutors into the proper position from which they can view the scene and recall the events that took place there. in a similar way, kerry’s use of the phrase “mission accomplished” invokes the narrative articulated by bush aboard the aircraft carrier. as with the phrase “weapons of mass destruction,” the phrase “mission accomplished” is embedded within an evaluative framework that reshapes its meaning. instead of indexing a valid set of truth claims, it now points to a set of incredulous claims. in turn, kerry positions himself as someone with a better handle on the ‘real’ truth; but this stance is made possible by first drawing upon the reservoir of words previously uttered by his opponent. in the final highlighted portion of this example, kerry uses the reported speech frame to typify bush’s prior discourse about his administration’s desire to streamline the military and wage war with a smaller, more nimble force. kerry reports the president to have said, “we could fight the war on the cheap.” this typifying speech (parmentier 1993, irvine 1996) emphasizes the content of bush’s prior discourse rather than its verbatim form. and importantly, the words used to convey this content imbue the message with an implicit evaluation. that is, through these reported words, kerry provides a preferred interpretation for how the discourse should be read. in particular, the phrase “on the cheap” conveys a negative evaluation of bush’s military policy. as bakhtin (1981) notes, prior words are “transmitted with varying degrees of precision and impartiality (or more precisely, partiality)” (330). for this reason, tannen (1989) prefers the term “constructed dialogue” to reported speech. the key point here is that reported speech frames provide an important means by which speakers reshape prior text, whether explicitly through accompanying metapragmatic commentary or implicitly through constructed dialogue that contains embedded evaluations. in excerpt (4), we see the political struggle over entextualization as political actors engage with their opponents’ words in an effort 7 hodges: the dialogic emergence of ‘truth’ in politics published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 8 to recontextualize them authoritatively and imbue them with their own preferred reading. 6. parody in the subversion of truth claims perhaps the most interesting aspect of intertextual connections is the way previously uttered phrases can be reanimated through parody. as bakhtin (1981) notes, “by manipulating the effects of context […] it is, for instance, very easy to make even the most serious utterance comical” (340). moreover, parody can be an effective tool of subversion. not only can it seriously challenge the truth claims of a political opponent, but in doing so, it can give play to an alternative narrative. i illustrate this with an example from an address given by rev. joseph lowery at the coretta scott king funeral in february 2006. with the current and past living presidents sitting behind him on the dais, lowery lifted the phrase “weapons of mass destruction” out of bush’s prior discourse, and reanimated it in his speech. part of the power of this example comes from the genre lowery chose: speaking in poetic verse. i have attempted to capture some of this verse by transcribing the example into lines and stanzas, as seen in (6). (note especially the underlined portions.) 6) from lowery’s speech at the coretta scott king funeral, february 7, 2006 she extended martin’s message against poverty, racism and war. she deplored the terror inflicted by our smart bombs on missions way afar. we know now there were no weapons of mass destruction over there ((23 sec cheers)) but coretta knew, and we knew, that there are weapons of misdirection right down here. millions without health insurance, poverty abounds, for war billions more, but no more for the poor. 8 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/1 doi: https://doi.org/10.25810/gkzf-q768 the dialogic emergence of ‘truth’ in politics 9 this example illustrates how genre can regulate intertextual relations. on the surface, the genre conventions associated with poetic verse allow for greater freedom in reanimating the words, especially for the incorporation of puns. as bakhtin (1986) notes, “speech genres in general submit fairly easily to reaccentuation, the sad can be made jocular and gay, but as a result something new is achieved” (87). the recontextualization in this example takes a serious issue and inserts it into a jocular frame that allows it to be transformed with a great deal of partiality. despite (or perhaps because of) the levity of the frame, the reaccentuation of the words seriously challenges the claims associated with “weapons of mass destruction” in bush’s narrative. instead of bolstering bush’s story, the pun on the phrase “weapons of mass destruction” undermines it; and it does so without an overly didactic tone. as a result, a serious point is made subversively. the effect of lowery’s incorporation of these words into his speech is to further a dialogue between alternative perspectives poised against one another in the politics of truth. these perspectives differ on the veracity and sincerity of the bush administration’s truth claims about the possession of weapons of mass destruction by saddam hussein. while the truth claim asserted by the bush administration gained powerful sway in public discourse prior to and immediately after the invasion of iraq, the opposing side in the debate has been compiling their own talking points to forward an alternative truth claim. this larger dialogue forms the backdrop to lowery’s address, even though it is a speech made by one person in what might traditionally be characterized as a monologue. as seen earlier, the phrase “weapons of mass destruction” carries indexical links to the narrative espoused by the bush administration. incorporation of this short phrase into the current context is sufficient to conjure up that larger text. in this way, as briggs and bauman (1992) note, “a crucial part of the process of constructing intertextual relations may be undertaken by the audience” (157). moreover, this reference to “weapons of mass destruction” and the subsequent play on those words—“weapons of misdirection”—reshapes the meaning of this key phrase in the national dialogue. basso’s work on apache moral narratives is useful to further explore lowery’s speech. recall how the apache invoke place-names in conversation to conjure up an entire story associated with that place. from that invoked story, a moral is drawn to be applied to the current purposes of the situation (cf. hanks 1989: 116). in particular, the moral is aimed at a specific individual who is present; and as the apache describe, the “stories go to work on you like arrows” (basso 1996: 38). briggs and bauman explain it this way: the point of the performance [in apache place narratives] is to induce an individual who is present to link her or his recent behavior—and what community members are saying about it—to the moral transgression committed in the story. (briggs and bauman 1992: 157) 9 hodges: the dialogic emergence of ‘truth’ in politics published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 10 the effect of the lowery example is somewhat similar. with bush sitting behind lowery on the dais, lowery (as the apache would metaphorically say) “shoots an arrow” at him by incorporating the phrase “weapons of mass destruction” into his verse. in reanimating these words, lowery turns them against bush as a reminder of the cumulative evidence against his administration’s truth claims. in effect, this works as a reminder directed at bush of his administration’s moral transgression and what critics are saying about it.6 lowery thus lodges a rebuttal as part of the larger dialogic struggle over truth in american politics. moreover, the words spoken by lowery in this particular context give play to a narrative in opposition to the one told by bush. importantly, this sound bite from lowery’s speech was itself subsequently recontextualized in the media in the weeks that followed; and lowery made appearances on fox news as part of the continuing discursive competition over the recontextualization of the phrase “weapons of mass destruction.” 7. conclusion the different discourse excerpts explored earlier form part of a larger national dialogue. in looking at the recycling of key phrases across different contexts, researchers gain a snapshot of the way intertextual series are drawn upon by political actors to engage in this dialogue and produce differing truth claims. as social actors draw from this reservoir of prior words, they work to reshape the larger social meanings associated with those words. in short, the process of entextualization is a political act in that lifting words and voices out of a prior context and recontextualizing them in a different setting imbues them with new interpretations. thus, truth in political discourse should not merely be analyzed as a product of the individual style of the politician to persuade or deceive, but as the confluence of various texts and discourses—as emergent from a dialogic process. political discourse is, like bakhtin (1981) says of the novel, “a system of languages that mutually and ideologically interanimate each other” (47). the effectiveness of rhetoric, therefore, comes from the interpretive web into which it enters. 6 this is also a prime example of signifying. that is, as mitchell-kernan (1972) describes, it is “a way of encoding messages or meanings in natural conversations which involves, in most cases, an element of indirection. this kind of signifying might be best viewed as an alternative message form, selected for its artistic merit, and may occur embedded in a variety of discourse” (165). 10 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/1 doi: https://doi.org/10.25810/gkzf-q768 the dialogic emergence of ‘truth’ in politics 11 references álvarez-cáccamo, celso. 1996. “the power of reflexive language(s): code displacement in reported speech.” journal of pragmatics 25: 33-59. austin, j.l. 1999 [1962]. “how to do things with words.” in the discourse reader, adam jaworski and nikolas coupland (eds.), 63-75. new york: routledge. bakhtin, mikhail. 1981. the dialogic imagination, caryl emerson and michael holquist (trans.), michael holquist (ed.). austin: university of austin press. bakhtin, mikhail. 1986. speech genres and other late essays, vern w. mcgee (trans.), caryl emerson and michael holquist (eds.). austin: university of austin press. basso, keith. 1996. wisdom sits in places: landscape and language among the western apache. albuquerque: university of new mexico press. bauman, richard and briggs, charles l. 1990. “poetics and performance as critical perspectives on language and social life.” annual review of anthropology 19: 59-88. bauman, richard. 2005. “commentary: indirect indexicality, identity, performance: dialogic observations.” journal of linguistic anthropology 15(1): 145-150. blommaert, jan. 2005. discourse: a critical introduction. cambridge: cambridge university press. briggs, charles l. and bauman, richard. 1992. “genre, intertextuality, and social power.” journal of linguistic anthropology 2(2): 131-172. bucholtz, mary and hall, kira. 2004. “language and identity.” in a companion to linguistic anthropology, alessandro duranti (ed.), 369-394. malden, ma: blackwell. buttny, richard. 1997. “reported speech in talking race on campus.” human communication research 23(4): 477-506. buttny, richard. 1998. “putting prior talk into context: reported speech and the reporting context.” research on language and social interaction 31(1): 45-58. duranti, alessandro. 1993. “truth and intentionality: an ethnographic critique.” cultural anthropology 8(2): 214-245. gal, susan. 2006. “linguistic anthropology.” in encyclopedia of language and linguistics, keith brown (ed.), 171-185. hanks, william f. 1986. “authenticity and ambivalence in the text: a colonial maya case.” american ethnologist 13(4): 721-744. hanks, william f. 1989. “text and textuality.” annual review of anthropology 18: 95-127. hill, jane. 2005. “intertextuality as source and evidence for indirect indexical meanings.” journal of linguistic anthropology 15(1): 113-124. hodges, adam. 2007. “the narrative construction of identity: the adequation of saddam hussein and osama bin laden in the ‘war on terror.’” in 11 hodges: the dialogic emergence of ‘truth’ in politics published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 12 discourse, war and terrorism, adam hodges and chad nilep (eds.), 67-88. amsterdam: john benjamins. hymes, dell. 1974. foundations in sociolinguistics. philadelphia: university of pennsylvania press. irvine, judith t. 1996. “shadow conversations: the indeterminacy of participant roles.” in natural histories of discourse, michael silverstein and greg urban (eds.), 131-159. chicago: university of chicago press. jakobson, roman. 1960. “closing statement: linguistics and poetics.” in style in language, thomas sebeok (ed.), 360-377. cambridge: mit press. jervis, robert. 2006. “understanding beliefs.” political psychology 27(5): 641663. kristeva, julia. 1980. “word, dialogue, and novel.” in desire in language: a semiotic approach to literature and art, leon s. roudiez (ed.), 64-91. new york: columbia university press. mannheim, bruce and tedlock, dennis. 1995. “introduction.” in the dialogic emergence of culture, dennis tedlock and bruce mannheim (eds.), 1-32. urbana and chicago: university of illinois press. mitchell-kernan, claudia. 1972. “signifying and marking: two afro-american speech acts.” in directions in sociolinguistics, john j. gumperz and dell hymes (eds.), 161-179. new york: holt, rinehart and winston. ochs, elinor. 1992. “indexing gender.” in rethinking context: language as interactive phenomenon, alessandro duranti and charles goodwin (eds.), 335-358. cambridge: cambridge university press. parmentier, richard. 1993. “the political function of reported speech: a belauan example.” in reflexive language: reported speech and metapragmatics, john lucy (ed.), 261-286. cambridge: cambridge university press. sacks, harvey. 1992. lectures on conversation. cambridge: blackwell. silverstein, michael. 1976. “shifters, linguistic categories, and cultural description.” in meaning in anthropology, keith basso and henry selby (eds.), 11-55. albuquerque: university of new mexico press. silverstein, michael. 1985. “language and the culture of gender: at the intersection of structure, usage, and ideology.” in semiotic mediation: sociocultural and psychological perspectives, elizabeth mertz and richard j. parmentier (eds.), 219-259. silverstein, michael. 2003. “indexical order and the dialectics of sociolinguistic life.” language and communication 23: 193-229. tannen, deborah. 1989. talking voices: repetition, dialogue, and imagery in conversational discourse. new york: cambridge university press. voloshinov, v.n. 1971. “reported speech.” in readings in russian poetics: formalist and structuralist views, ladislav matejka and krystyna promorska (eds.), 149-175. cambridge, ma: mit press. voloshinov, v.n. 1973. marxism and the philosophy of language, ladislav matejka and i.r. titunik (trans.). new york: seminar press. 12 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/1 doi: https://doi.org/10.25810/gkzf-q768 colorado research in linguistics 6-2008 the dialogic emergence of ‘truth’ in politics: reproduction and subversion of the ‘war on terror’ discourse adam hodges recommended citation microsoft word cril_paper_hodges.doc 1. introduction the modification patterns of “compound indefinite pronouns” in english like something, everybody, and nowhere (quirk et al. 1985, cf. wu 2021) are intriguing because unlike other pronouns, this word class can take adjectival modifiers, but unlike nouns, these modifiers cannot occur pre-(pro)nominally (quirk et al. 1985, kishimoto 2000, larson & marušič 2004, leu 2004, wu 2021). for example, something interesting is grammatical but *interesting something is not. 1 many theorists have explained indefinite pronoun modification by comparing it to nominal adjectival modification. the two main nominal adjectival modification constructions in english are the attributive construction, in which adjectives come before a noun, and the predicative construction, in which adjectives come after a predicate that relates the adjective to a noun (bolinger 1967). however, there are a limited set of scenarios in which an adjective can occur in a postpositive position, directly after a noun, in english. first, adjectives with complements or coordination can come after nouns, as in an actor suitable for the part and soldiers timid or cowardly (quirk et al. 1985). second, a specific class of attributive adjectives are able to follow nouns without complementation, construing a relationship with the noun that is temporary or “stage-level” (carlson 1977): visible, navigable, responsible, and other -ible/-able adjectives, ablaze, afloat, and other aadjectives, as well as adjectives such as present, concerned, and involved (bolinger 1967, quirk et al. 1985, larson & marušič 2004). the existence of what i call the post-indefinite pronoun modification construction (pipm) has stumped many proponents of the generative syntactic perspective: how, if at all, are examples like something interesting related to post-nominal modification? from what underlying structures do such examples originate? do they involve attributive adjectives, predicative adjectives, postpositive adjectives or something else entirely? various analyses have been proposed to explain post-indefinite pronoun modification. kishimoto (2000) argues for a movement-based analysis of these structures, in which everything interesting is a transformation of the np deep structure every interesting thing. larson and marušič (2004) critique this analysis, articulating that the adjectives in this construction must originate “in place” because of certain behavioral similarities between postnominal adjectival modification and postindefinite pronoun adjectival modification. for example, like all the stars visible, everything visible has a stage-level (temporary/episodic) interpretation, while prenominal modification is ambiguous between stage-level (temporary/episodic) and individual-level (inherent/intrinsic) interpretations. however, larson and marušič’s (2004) arguments rely on comparison to a different set of postpositive adjectives (e.g. visible, navigable) than those that are most typical of post-indefinite pronoun modification (e.g. interesting, unusual). wu (2021) adds on to larson and marušič’s analysis, arguing that adjectives like unusual and interesting are “coerced” into postpositive position, while visible and navigable are inherently postpositive. however, like larson & marušič (2004), wu’s analysis assumes that all types of post-modification, whether nominal or pro-nominal, have the same types of interpretations. i suggest that there are key differences between the semantic construal associated with postnominal modification and pipm that set pipm apart. in this paper, i draw on the framework of construction grammar (cxg; michaelis 2012; hilpert 2014; goldberg 1995, 2006) to demonstrate that post-indefinite pronoun modification is best understood as a separate construction from postnominal modification. although it is true that post-indefinite pronoun modification shares various qualities with postnominal modification, tokens like everything interesting have individual-level construal rather than stage-level construal. for example, while everything visible means everything that is currently visible, everything interesting means everything that is inherently interesting. this difference is evidence of a separate schematic form-meaning pairing from other cases of postpositive modification. in pipm, gradable evaluative adjectives follow indefinite pronouns, and pick an indefinite entity or event out of a class of entities or events prototypically socially evaluated as the adjective in question for the situation being described. for example, something weird happened construes a set of prototypically weird events that can happen and identifies the referent as one of these events. this paper will be structured as follows: in §2, i will give several examples of the pipm construction. in §3, i will review previous analyses of this construction that compare postindefinite pronoun modification to postnominal modification, involving what i call “restricted access” adjectives. in §4, i will present preliminary semantic arguments for why pipm should be understood as a construction separate from not only attributive and predicative modification, but also from postnominal modification involving restricted access adjectives. i also discuss why the framework of cxg is perfectly suited to describe the unique construal associated with pipm. in §5, i will follow up with additional syntactic evidence from the contemporary corpus of american english (coca) 2 that pipm is a separate construction from postnominal modification. the remaining sections of the paper will outline the form and meaning of pipm (§6) and demonstrate some coercion effects (§7), before concluding remarks in §8. 2. examples of pipm in this section i provide several examples of pipm from coca. this construction consists of any compound indefinite pronoun, and any gradable, evaluative adjective. compound indefinite pronouns are those made up of two morphemes, a determiner/quantifier morpheme (e.g. every-, some-, any-, no-) and a nominal morpheme (e.g. -one, -body, -thing), and thus do not include indefinite pronouns like some and one (quirk et al. 1985). based on the corpus analysis i conduct in §5, common evaluative adjectives in this construction are as follows: new, wrong, different, unusual, cool, good, bad, suspicious, strange, interesting, funny, special, nice, stupid, similar, negative, terrible, unexpected, & amazing. see examples of pipm in 1: 1) a. and that’s where we see something new and potentially something very promising b. do you see anything wrong in that? c. they are willing to spend more to get something special. d. he saw nothing unusual at first e. everything bad was over f. the deceased must have been somebody important g. just find someplace nice h. one thing good about late spring is… while i agree with larson & marušič (2004) that transformational analyses that involve movement of thing from post-adjectival to pre-adjectival position are not appropriate, thing is still relevant to this construction. when preceded by a quantifier, thing and other semantically “light” nouns play a similar semantic role as compound indefinite pronouns and therefore fit into this pattern (see 1h). 3. previous analyses kishimoto’s (2000:558-559) analysis outlines the problem of post-indefinite pronoun modification with a set of contrasts between attributive adjectival modification of nouns vs. indefinite pronouns, reproduced below in examples 2-7. it should be noted that these adjectives must precede nouns (as in 2a, 3a, & 4a) and must follow indefinite pronouns (as in 5b, 6b, & 7b): 2) a. every interesting book b. *every book interesting 3) a. a delicious dish b. *a dish delicious 4) a. cold rooms b. *rooms cold 5) a. *interesting everything b. everything interesting 6) a. *delicious something b. something delicious 7) a. *cold someplace b. someplace cold to account for these distinctions, kishimoto (2000) proposes an n-raising analysis in which the “light nouns” thing and place can move from a location in an np, after the adjective, to a position in a num phrase, before the adjective, as outlined in 8a (pre-movement) and 8b (postmovement): 8) a. [dp every [nump [np interesting thing]]] b. [dp every [nump thing [np interesting________]]] larson & marušič (2004) argue that the movement analysis cannot be correct, by drawing comparisons between post-indefinite pronoun modification and the behavior of certain postnominal adjectives, including visible, navigable, and responsible. they build off of bolinger’s (1967) argument against a movement analysis that ties attributive and postnominal uses of these adjectives to the same underlying structure. bolinger instead analyzes postnominal adjectives as reduced relative (predicative) clauses, as in 9b, reduced from 9a: 9) a. the stars that are visible b. the stars visible although larson & marušič (2004) provide many reasons why a movement analysis of postindefinite pronoun modification cannot be correct, here i will review only two of them: 1) attributive-only adjectives do not occur with indefinite pronouns and 2) indefinite pronoun modification involves stage-level construal, like postnominal modification. firstly, post-indefinite pronoun modification cannot be underlyingly attributive because indefinite pronouns cannot be modified by the adjectives that bolinger (1967) identifies as only occurring attributively. bolinger (1967) points out that there are specific adjectives, such as live and mere, that occur attributively, as in 10, but do not occur predicatively or postnominally, as in 11 and 12 respectively. larson & marušič (2004) points out that these attributive-only adjectives also do not occur following indefinite pronouns, as in 13 (larson & marušič 2004:273, wu 202:826): 10) a. live animal b. mere idea 11) a. *this animal is live (cf. this animal is alive) b. *no idea is mere 12) a. *an animal live b. *no idea mere 13) a. *something live (cf. something alive) b. *nothing mere therefore, larson & marušič (2004) suggest that post-indefinite pronoun modification cannot be underlyingly attributive, otherwise indefinite pronouns would be able to be post-modified by attributive-only adjectives. secondly, and crucial to the current analysis, larson & marušič (2004) draw on bolinger’s (1967) observation that while attributive modification can either be interpreted as stage-level (temporary/episodic) or individual-level (inherent), postnominal modification necessarily has a stage-level construal. larson & marušič (2004) argue that like postnominal modification, indefinite pronoun modification also only has a stage-level construal. larson & marušič (2004:274) contrast attributive modification in 15a & 16a with postnominal modification in 15b & 16b: 15) a. list all the visible stars, whether we can see them or not. b. ??list all the stars visible, whether we can see them or not. 16) a. list all the responsible individuals, whether they were involved or not. b. ??list all the individuals responsible, whether they were involved or not. attributive modification, in 15a, can refer to stars that are in general visible to the naked eye – or “inherently” visible (individual-level) – or it can refer to stars that are currently (temporarily/episodically) visible (stage-level). because visible stars can have both readings, it can occur with a continuation that directly references the two possibilities of currently visible or currently invisible. similarly, in 16a, responsible individuals can be interpreted either as individuals who are responsible in general, in other words, trustworthy (individual-level), or responsible for a specific situation (stage-level), and thus can occur with a similar continuation that denies the possibility of current involvement. nominal post-modification on the other hand, does not have both possibilities. stars visible in 15b necessarily means the stars that are temporarily currently visible (stage-level), and thus the same continuation sounds odd (denoted by ??). similarly, individuals responsible necessarily refers to individuals responsible for some current situation (stage-level). larson & marušič’s (2004) then compare these construals to 17, examples of post-indefinite pronoun modification: 17) a. ??list everything visible, whether we can see it or not. b. ??list everyone responsible, whether they were involved or not. like 15b, everything visible in 17a necessarily means everything that is currently visible (stagelevel) and cannot mean everything that is visible in general (individual-level), and like 16b, everyone responsible in 17b necessarily means someone who is responsible for some current act (stage-level). therefore, both of these cannot take the continuation that contrasts two possibilities. larson & marušič (2004) propose that if everything visible was underlyingly every visible thing, that it would have the same semantic construal as the attributive modification pattern in 15a & 16a – that it would be ambiguous between stage-level and individual-level. they thus provide convincing evidence that post-indefinite pronoun modification does not originate prenominally (attributively). however, while larson & marušič (2004) begin their discussion by citing kishimoto’s examples (2-7) of evaluative adjectives that can occur after indefinite pronouns but not after nouns, their argumentation relies solely on comparison to adjectives that can occur after nouns, the -ible/-able adjectives, such as visible and responsible. many of these adjectives convey that there is restricted access to the noun they modify, in other words, that only a limited number of the noun is visible or responsible (and that others are non-visible, or not responsible). therefore, i’ll call this group of adjectives “restricted access” (ra) adjectives, for ease of reference.3 this is in general, a different set of adjectives than those that can only follow indefinite pronouns and which typically occur with pipm (see §2), which i’ll refer to as pipm adjectives. larson & marušič (2004:270) admit that their account is not able to explain, if post-indefinite pronoun adjectives “originate postnominally” in a similar fashion to postnominal modification, what prevents *every book interesting, *a dish delicious, and *rooms cold in 2-4. an analysis of the pipm pattern must address the distinction between adjectives that can generally follow nouns (ra adjectives) vs those that cannot (pipm adjectives), and cannot rely upon the syntax and semantics of one modification pattern to explain the other. along these lines, wu (2021) expands upon larson & marušič’s (2004) analysis, proposing a syntactic explanation for why *every book interesting doesn’t occur but something interesting does. while wu (2021) agrees that tokens like stars visible are reduced from relative clauses (stars that are visible), he claims that only certain adjectives (ra adjectives) are inherently postpositive and thus can undergo this reduction. adjectives that typically (except for in pipm) occur prenominally (e.g., interesting) are instead coerced to postpositive position specifically when occurring with indefinite pronouns (not nouns). wu (2021) argues that this coercion process occurs because the “prenominal” modifier position is not available, as the “determiner” (e.g., every) and “noun” (e.g., thing) pieces of compound indefinite pronouns cannot be broken up by the insertion of a prenominal modifier. therefore, the modifier needs to occur after the compound indefinite pronoun, since modifiers cannot occur before determiners. this explanation is favorable because it focuses on the unique morphosyntactic properties of indefinite pronouns – on their properties that fall between phrase-hood and word-hood. however, wu (2021:836) claims that as a corollary of this coercion process, “the placement of potential attributive adjectives in postposition will restrict them [to] ‘temporariness.’” in other words, he again explains the semantics of pipm modification in terms of ra modification, even while successfully separating the syntactic patterns involved. while this construal effect is often true for ra adjectives like visible, it is not true for pipm adjectives. i will demonstrate evidence for this assertion in the next section. 4. pipm: a construction in its own right some theorists have suggested that post-indefinite pronoun modification is underlyingly attributive and undergoes movement to occur in postpositive position (kishimoto 2000). others have suggested that it is underlyingly postpositive, possibly reduced from a predicative relative clause (larson & marušič 2004), or coerced to postpositive position because of the specific morphosyntactic structure of compound indefinite pronouns (wu 2021). in this section, i review the evidence for each position, detailing in what ways post-indefinite pronoun modification is similar to attributive modification, and in what ways it is similar to ra postnominal and predicative modification. ultimately, i argue that post-indefinite pronoun modification shares particular features with each of these types of modification and is therefore best described as a separate construction in its own terms. as discussed in the last section, previous analyses have demonstrated that post-indefinite pronoun modification is similar to postnominal and predicative modification patterns in that attributive-only adjectives such as live and mere cannot modify indefinite pronouns. pipm is clearly not the same modification pattern as attributive modification. previous approaches (larson & marušič 2004, wu 2021) have also argued that post-indefinite pronoun modification produces a temporary (stage-level) construal. however, i argue that this is often true for ra adjectives but not for pipm adjectives. for example, in 18a, an example with a pipm adjective, goodness is characteristic of everything (individual-level), not a temporary quality. similarly, in 18b (repeated from 1f), importance is a characteristic quality of the deceased (individual-level), that does not disappear after their death. lastly in 18c (repeated from 1d), unusualness is a characteristic quality of the possible entities or events to be seen (individual-level). what is temporary in this case is the period in which one has not seen one of these items. 18) a. you remind me of everything good. b. the deceased must have been somebody important. c. he saw nothing unusual at first. due to the individual-level construal involved in pipm, in addition to this pattern showing conflict with attributive-only adjectives, it is in fact also rare with predicative-only adjectives, like afraid and sorry, that tend to convey temporary feelings rather than general characteristics. this is not predicted by larson & marušič’s (2004) and wu’s (2021) claims that all postpositive adjectives have a temporary construal. while the man is afraid is felicitous, someone afraid is distinctly odd. in coca, although there are 54 hits for indef-pronoun afraid, many of these involve secondary predicates or adjectival complements. there is only one true example of pipm, in 19: 19) mr-hamill: they -i learned that, that they -they dwell on fear. if you -if you give into them, you know, that -that -that gives them a big high. they want to see somebody afraid. and -and -and i really wasn't -didn't have to -to try to act like i wasn't. being afraid is typically not something that characterizes somebody, but in 19, hamill construes fright as a characteristic quality of someone the individuals in question are looking for. this is a marginal example; in general, afraid doesn’t tend to occur in pipm. similarly, there are no true examples of pipm with sorry. thus, while post-indefinite pronoun modification shares some features with predicative and postnominal modification (including conflict with attributive-only adjectives), it shares individual-level construal with attributive modification. this makes examples of pipm different from reduced relatives – someone who is sick doesn’t mean the same thing as someone sick, because someone sick construes sickness as an individual-level property, while the predicative relative clause construes it as a temporary stage-level one. this is very different from what we see with post-indefinite modification involving ra adjectives, like visible. as discussed in the previous section, something visible does mean the same thing as something that is visible. thus, pipm is a different constructional pattern than attributive, predicative, or postnominal modification, and should not be explained in terms of any of these other syntactic patterns. instead, the framework of construction grammar (cxg) can be utilized to explain the idiosyncrasies associated with post-indefinite pronoun modification. within cxg, syntactic patterns are assigned specific meanings and/or functions, just as words are (michaelis 2012, hilpert 2014, goldberg 1995, 2006). whereas generative perspectives find it difficult to explain cases in which the same schematic syntactic structure is associated with two different functions or meanings, cxg can account for such patterns. for example, goldberg (1995:204-209) identifies two different meanings associated with the “way” construction, a “means” version and a “manner” version labeled in 20: 20) a. means: in some cases, passengers tried to fight their way through smoke-chocked hallways to get back to their cabins to get their safety jackets. b. manner: …he was scowling his way along the fiction shelves in pursuit of a book. the means version of this construction involves interpretations of the verb (e.g. fight) in which the associated action is the means of creating a path, while the manner version involves interpretations of the verb (e.g. scowl) in which the associated action is an activity occurring at the same time as traversing a path. in the present analysis, we are dealing with a similar case, in which the same surface form, postpositive modification, is associated with two separate functions, restricted stage-level modification (ra modification), and evaluative individual-level modification (pipm). if there are two separate constructions associated with the same surface form, we should be able to find examples in which the same adjective is used in both constructions, with two different construals. indeed, these cases can be found. while bolinger (1967) says that “the man responsible is unambiguously ‘to blame’ and the responsible man is almost unambiguously ‘trustworthy’” (p. 4), and larson & marušič (2004:273) claim that post-indefinite examples always share the episodic (to blame) reading with postnominal modification, we do see post indefinite examples like those in 21, in which someone responsible means someone that is trustworthy: 21) a. that's great, " she says. " the landlord is definitely looking for someone responsible. " " i'm that person, " i say. b. if you could pair him with someone responsible maybe a girl. these examples, although they include ra adjectives, appear to be examples of the pipm construction, with individual-level construal, rather than examples of the typical pattern that bolinger (1967) and larson & marušič (2004) describe for postnominal ra modification, involving stage-level construal, as in 22: 22) i did nothing... not my fault. nobody says it is. take it easy, right? we will find someone responsible. thus, rather than post-indefinite pronoun modification following the phrase structure rules for nouns or requiring movement or coercion to produce, the modification pattern that occurs with indefinite pronouns is specific to this lexical class – it is a formal idiom (michaelis 2012), associated with its own form and function. in the next section, i aim to demonstrate that while there are examples of ra adjectives that occur in the pipm construction, in general, pipm adjectives and ra adjectives are separate sets of adjectives that occur in separate syntactic patterns. 5. corpus analysis in the previous section, i demonstrated that pipm and ra modification have different semantic construals. in the current section, i demonstrate that although their immediate surface structure is the same – they both occur postpositively – they occur in different larger syntactic patterns. specifically, pipm occurs in at least two larger syntactic patterns that ra adjectives are rare or do not occur in: after verbs of perception (e.g., see something unusual) and before verbs of occurrence (e.g., something unusual happened). the general goal of this corpus case study is to show that adjectives common in previous arguments that compare postnominal modification to post-indefinite pronoun modification – the ra adjectives – don’t occur in the same syntactic contexts as pipm adjectives. for ra adjectives, i selected a small group of adjectives that consistently have come up in the literature: visible, navigable, available, possible, and present (a non-ible adjective, that also participates in the same restricted access pattern). for pipm adjectives, i first selected the larger syntactic patterns that appear to be specific to pipm (after verbs of perception and before verbs of occurrence), in order to isolate examples of pipm, and then identified adjectives that occur most commonly in these patterns. this led to the group of adjectives similar, strange, terrible, unusual, different, and new. although the adjective wrong is also very frequent in pipm, it is often used as a secondary predicate or adverbial, and its use in pipm is therefore difficult to isolate. the need for an analysis involving larger syntactic structures becomes apparent when examining the results of an initial corpus search in table 1 below. many more nouns occur before ra adjectives than indefinite pronouns, but only some pipm adjectives have more indefinite pronouns than nouns preceding them. in order to show that pipm occurs in unique syntactic patterns, additional examination of these examples is needed. restricted access adjs indef-pronoun adj something visible noun adj man visible visible 102 2201 navigable 0 19 available 351 31845 possible 1210 9416 responsible 402 7538 present 236 11825 pipm adjs similar 2664 9486 strange 1278 676 terrible 1336 325 unusual 1848 275 different 6684 5291 new 14498 11489 table 1. restricted access adjs vs pipm adjs after nouns & indefinite pronouns (coca) in table 1, there are several types of examples of pipm adjectives following nouns that aren’t true examples of post-modification. for example, included in these numbers are secondary predicates as in 23a, adverbials as in 23b, and examples with adjectival complements as in 23c-d. examples with adjectival complements are not examples of pipm because all adjectives can occur postpositively if they occur with complements, as discussed in §1. 23) a. it's harder to make those jokes and make comedy funny if you don't have any profanity b. after completing the book i saw things different c. …creating a self funded program similar to those in at least 10 other states… d. …we'd found 400 species new to the park… thus, corpus work with this construction requires manual attention to weed out irrelevant examples. i thus examined larger syntactic patterns for two reasons: 1) i thought i would be more likely to isolate examples of pipm within these patterns, and 2) to create a smaller set of data to manually remove examples of secondary predicates, adverbials, adjectives with complements, as well as idioms like brand new. for both analyses below, i manually examined all tokens, in other words, examples with both ra and pipm adjectives, and those after both nouns and indefinite pronouns, removing all examples of these irrelevant patterns. although coordinated postposed adjectives are another pattern that license post-modification, many valid examples of pipm involve coordination (as in 1a). therefore, i retained these examples and discuss an example of this pattern below. 5.1 verbs of perception table 2 shows the results of searches in coca for patterns involving post-modified nouns and post-modified indefinite pronouns with ra & pipm adjectives, following the perception verbs see, hear, taste, smell, and touch. two separate noun searches were needed to account for singular & plural tokens. during manual review of the data, in addition to the types of examples discussed above, i also removed tokens in which the constituency wasn’t clear. for example, in 24, a wh-question constituency test is awkward: 24) a. i don’t see anything unusual about this. b. ??what do you see _____ about this? restricted access adjs see/hear/taste/smell/touch indef-prn adj see something visible see/hear/taste/smell/touch noun adj see men visible see/hear/taste/smell/touch * noun adj see a man visible visible 1 0 0 navigable 0 0 0 available 2 2 16 possible 1 0 4 responsible 0 0 1 present 0 0 2 pipm adjs similar 73 1 0 strange 119 1 0 terrible 22 0 0 unusual 187 0 0 different 191 0 0 new 276 0 0 table 2: restricted access adjs vs pipm adjs in the “perception-verb x adj” pattern (coca) as expected, indefinite pronouns with pipm adjectives are much more common after perception verbs than both indefinite pronouns with ra adjectives or nouns with either type of adjective. examples of pipm adjectives following indefinite pronouns in this pattern are shown below in 25: 25) a. did you hear anything unusual last night? b. because whenever he saw something new and interesting, or new and ridiculous, he always wondered what she'd have to say about it. while ra adjectives following indefinite pronouns are less common (at least after perception verbs), they are still grammatical, as the examples in larson & marušič (2004) show. one example of this type is given in 26. this example has the stage-level construal predicted in previous analyses, while pipm examples in 25 are individual-level. 26) we need a lefty in the outfield and i don't see anyone available. lastly, while it is expected that ra adjectives would also follow nouns, it is unexpected that pipm adjectives would follow nouns. therefore, these tokens, in 27, require further discussion: 27) a. we saw visions strange and foreboding, but we kept them to ourselves, because heaven blesses the meek b. but just hearing and seeing videos similar on the internet, it just made me uncomfortable. 27a is an example that involves coordination, and thus is a predictable example of postmodification (§1). however, 27b is not easily explainable. while “on the internet” is a secondary predicate of see rather than a complement of similar, perhaps this example may also be licensed due to the “heaviness” of this clause (bolinger 1967, larson & marušič 2004). overall, such examples appear to be marginal. examples of pipm are common after verbs of perception, while examples of restricted access modification are less common after verbs of perception. crucially, the adjectives identified as pipm adjectives only felicitously occur after indefinite pronouns. 5.2 verbs of occurrence table 3 shows the results of searches in coca for patterns involving post-modified nouns and post-modified indefinite pronouns with ra & pipm adjectives, preceding the occurrence verbs happen, occur, begin, start, and be going on. restricted access adjs indef-pronoun adj happen/occur/start/begin/be going on something visible happened noun adj happen/occur/start/begin/be going on thing visible happened visible 2 0 navigable 0 0 available 0 1 possible 0 0 responsible 0 0 present 0 3 pipm adjs similar 140 0 strange 183 2 terrible 186 0 unusual 106 1 different 36 0 new 37 0 table 3. restricted access adjs vs pipm adjs in the “x adj occurrence-verb” pattern (coca) as in the previous section, it is clear that indefinite pronouns with pipm adjectives are more common preceding verbs of occurrence than both indefinite pronouns with ra adjectives or nouns with either type of adjective. examples of pipm adjectives following indefinite pronouns in this pattern are shown below in 28: 28) a. did anything strange happen when you were living there? b. it becomes more apparent that something terrible is going on inside kosovo. there are three postnominal examples of pipm with verbs of occurrence, two examples involving the same phrase (shown in 29a) from different parts of the movie killer tomatoes eat france, and one additional example in 29b: 29) a. when these four things strange occur as one... " the true king of france shall return with the sun. b. …that he will go away satisfied and not report back to the authorities that some thing unusual is going on in that household 29a is an example of quantified “light noun” thing, therefore an example of pipm. since this example also involves a rhyme, this suggests that the ordering of this token is creative and playful. 29b is also an example of pipm with a space between the “determiner” and “noun” parts of the compound indefinite pronoun. therefore, while these examples are marginal, they are explainable. in this short corpus case study, i have shown that despite the same surface structure of postmodification, there are larger syntactic contexts in which pipm is relatively common but restricted access modification is rare. by isolating pipm examples after verbs of perception and before verbs of occurrence, i have provided further evidence that pipm is a separate modification construction than restricted access modification, adding on to the semantic evidence in the previous section. the pipm adjectives that i examined rarely post-modify nouns in these patterns, and when they do, they are typically examples of the pipm pattern, with the quantified light noun thing or with a space between parts of the indefinite pronoun. furthermore, despite comparisons to ra adjectives in the literature on the pipm pattern, in the syntactic patterns analyzed here, ra adjectives rarely occur. 6. constructional analysis of pipm pipm is a formal idiom, or construction, because its form cannot be licensed by traditional phrase structure rules, it has a special interpretation, and because there are semantic constraints on the words that can occur in the construction (michaelis 2012). the formal structure of the pipm construction is as follows: [compound indefinite pronoun | quantifier thing(s)] (degree or essence adverb) evaluative adjective the meaning of the construction is a gestalt construal of an indefinite entity or event that is selected from a backgrounded larger category of entities or events, evaluated by prototypical societal norms as falling somewhere along the scale of the adjective in the typified context of use. context is important to this construction because something unusual means something different in say something unusual, taste something unusual, and something unusual happened: a different larger category of actions or entities (unusual things that are said, things that taste unusual, unusual things that happen) is construed for these different phrases. therefore, the larger category from which an indefinite pronoun is selected can be abstract or concrete, and made of entities or events, since indefinite pronouns can refer to many different things. the fact that similar and different are common in this construction emphasizes the social situatedeness and intersubjectivity of this construction, as social actors commonly employ it to discuss situations that are similar or different from what they are currently jointly focused on. however, despite a contextual construal, the categories evoked by this construction are contingent on conventional ideologies and social stereotypes about what kinds of things can be described as the adjective in question (or can be evaluated as similar or different from the entity in question). thus, while the adjectives in this construction are subjective and gradable, the construction identifies a class of entities or events that rely on conventional societal norms to establish what is terrible, unusual, strange, new, similar, or different in a given context. the construction doesn’t just identify one such thing, but a whole category of things, which emphasizes socially sanctioned or “typified” understandings of social action (gal & irvine 2019). pipm’s referent is a gestalt – rather than emphasizing the referents’ indefiniteness or the quality of the adjective, the construction construes both as equally important to its meaning. this distinguishes something unusual from similar ways of to say the same thing like something that is unusual, or some unusual thing, which have far fewer tokens in coca. as discussed in §3, while larson & marušič (2004) & wu (2021) argue that postnominal and post-indefinite modification construes adjectives as only temporarily or episodically modifying the nouns they describe, the kind of social meaning that pipm imparts is necessarily individuallevel, since it deals with stereotypes of social action. this can be shown by the fact that you can modify pipm with “essence” adverbs, such as fundamentally, essentially, and inherently, as in 30. it does not appear that these adverbs can be used in ra modification. 30) this is an exciting result that suggests something fundamentally different about what processes play a key role in the generation of mercury's magnetic field…. degree adverbs can also modify the adjectives in this construction. this includes adverbs such as so, very, slightly, and totally, which push the category in one direction on the scale construed by the adjective. finally, comparatives also are sanctioned, since this merely construes a category that is compared to another category. but superlatives, like ??something most different, are not attested in coca, since superlatives refer to the highest end of a scale rather than a categorical class associated with a region on a scale. something should be said about the effect of different indefinite pronouns in this construction, since the construction has a slightly different meaning with indefinite pronouns that start with the quantifiers no, some, any, and every. in these different cases, the backgrounded category against which the indefinite meaning is construed is the same, but the foregrounded part of the category, the referent of the overall construction, is different. with indefinite pronouns that start with no, as in the phrase nothing new, a category is construed of things prototypically socially evaluated as new, and the referent of the phrase is associated with none of these things. with some, we have seen that the referent is identified as an indefinite thing out of the larger category. any acts similarly to some, but is often used in negative contexts, questions and subjunctives. lastly, when the indefinite pronoun begins with every, the referent of the construction is the entire construed category. 7. coercion effects and social meaning one of the demonstrations of a meaningful and productive constructional pattern is the observation that its constructional meaning can coerce a new meaning from a lexical item that doesn’t match its selectional restrictions, in other words, the types of lexical items with which it normally combines (michaelis 2012). as i discussed in §6, pipm generally occurs with gradable, subjective adjectives that are associated with socially proscribed categories in a particular context. however, occasionally, non-subjective adjectives (31a), non-gradable domain adjectives (sullivan 2013; 31b-c), and proper nouns or adjectives formed from proper nouns (31d-g) can occur in pipm: 31) a. i was waiting for him to say something drunk like this baby needs us. this baby didn’t need us. b. nothing sexual happened beside a few erotic kisses. c. does that mean a christian may not say anything christian in an islamic state? d. and they stopped paying sarah palin to come into the office once every three months and say something sarah paliny e. when i get home from work, i build a fire and chuckle to myself that mr. murphy saw something thoreauvian in my nature f. i almost wanted to do a vegas theme for this but alas nothing vegas happened on this day in hockey history. g. and now for something completely… obama (gonzálvez-garcía 2014:281) in these examples, adjectives (or nouns) that typically are not considered gradable, subjective, or associated with social evaluation receive a construal as such. in 31a, an adjective that is not typically used to describe a subjective quality, drunk, is construed as a subjective evaluative term that describes a category of things one might say while drunk. in 31b-c, non-gradable adjectives that normally designate a domain under which types of activities or actions can be categorized, such as sexual and christian, are also construed as socially evaluated gradable categories. crucially, these tokens rely on shared social knowledge of what types of things can be evaluated as sexual or christian. lastly, there are several examples of pipm tokens that involve proper nouns or adjectival forms of proper nouns, in 31d-g. these examples rely on metonymic inferencing (gonzálvez-garcía 2014, danygier 2011) to evoke the quality associated with a category of socially evaluated entities, and thus presuppose expertise with the domain associated with the proper noun. since expertise often signals ones’ engagement with particular activities or communities, this construction may be leveraged in the construction of social meaning (eckert 2008, silverstein 2006). for example, something thoreauvian indexes one as both elite (possessing academic knowledge) and as a lover of nature. 8. conclusion in this case study i have shown that the post-indefinite pronoun modification, or pipm, construction is a separate adjectival construction than postnominal modification with restricted access adjectives. while the “light” noun thing can participate in this construction, other nouns cannot. in addition, while pipm involves individual-level construal, postnominal modification involves stage-level construal. thus, pipm should be analyzed on its own terms, and not as a form of postnominal modification as generative analyses have done (larson & marušič 2004, wu 2021). using the framework of construction grammar (michaelis 2012), i have shown that this construction has a specific form and semantic interpretation that cannot be predicted from its parts. its meaning plays into and (re)constructs ideologies about particular social qualities, and it even coerces adjectives and nouns that are not usually evaluative to evaluate entities and events along social scales. analysis of this construction can thus complement work at the intersection of syntactic analysis and the construction of social meaning (moore 2009). references bolinger, dwight. 1967. adjectives in english: attribution and predication. lingua 18.1–34. carlson, greg. 1977. a unified analysis of the english bare plural. linguistics and philosophy 1.413–456. dancygier, barbara. 2011. modification and constructional blends in the use of proper names. constructions and frames, 3.208–235. eckert, penelope. 2008. variation and the indexical field. journal of sociolinguistics, 12.453– 476. gal, susan and judith irvine. 2019. signs of difference: language and ideology in social life. cambridge university press. goldberg, adele. 1995. constructions: a construction grammar approach to argument structure. the university of chicago press. goldberg, adele. 2006. constructions at work: the nature of generalization in language. oxford university press. gonzálvez-garcía, francisco. 2014. “that’s so a construction!”: some reflections on innovative uses of “so” in present-day english. studies in functional and structural linguistics, ed. by maría de los ángeles gómez, francisco ruiz de mendoza ibáñez, and francisco gonzálvez-garcía, 271–294 (vol. 68). john benjamins publishing company. hilpert, martin. 2019. construction grammar and its application to english (2nd ed.). edinburgh university press. kishimoto, hideki. 2000. indefinite pronouns and overt n-raising. linguistic inquiry, 31.557– 566. larson, richard and franc marušič. 2004. on indefinite pronoun structures with aps: reply to kishimoto. linguistic inquiry, 35.268–287. leu, thomas. 2005. something invisible in english. university of pennsylvania working papers in linguistics, 11.143–154. michaelis, laura. 2012. making the case for construction grammar. sign-based construction grammar, ed. by hans boas and ivan sag, 31-67. csli publications. moore, emma and robert podesva. 2009. style, indexicality, and the social meaning of tag questions. language in society, 38. 447–485. quirk, randolph; sidney greenbaum; geoffrey leech; and jan svartvik. 1985. a comprehensive grammar of the english language. longman. silverstein, michael. 2006. old wine, new ethnographic lexicography. annual review of anthropology, 35.481–496. sullivan, karen. 2013. frames and constructions in metaphoric language. john benjamins publishing company. wu, zhen. 2021. compound pronouns in english. english language and linguistics, 25.825– 849. end notes 1 wu (2021:845) explains an exception that some indefinite pronouns can occur with articles and prenominal adjectives, however then they are construed as nouns, as in a very special someone. 2 all examples cited in this paper are from coca. 3 while it is likely the case that these postnominal adjectives occur in a particular restricted access construction, i stay agnostic on the constructional status associated with this postnominal modification since it is not the focus of this paper. developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses colorado research in linguistics. june 2010. vol. 22. boulder: university of colorado. © 2010 by cecily jill duffield developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibokonibo relative clauses* cecily jill duffield university of colorado subject-verb agreement is assumed to be the marking of the verb in an utterance as determined by properties of the subject. psycholinguistic models of agreement in speech production differ as to whether they treat this phenomenon as driven primarily by syntactic processes or semantic influences. but these models are based primarily on research in indo-european languages. this paper suggests that a useful approach to investigating the psycholinguistic mechanisms behind agreement in speech production is to extend the research to more typologically variant languages and more complex structures. relative clause data from a panoan language, shipibo-konibo, based on the work of valenzuela (2002) is presented here as an ideal case study for psycholinguistic research on syntactic and semantic influences on subject-verb agreement. shipibo-konibo has a flexible word order, and a variety of relative clause types and relativization strategies that display subjects and verbs in various positional relationships. the data is presented in the context of two psycholinguistic models of agreement production: the marking and morphing model (eberhard, cutting & bock 2005) and the maximal input model (vigliocco & hartsuiker 2002). 1. introduction the study of sentence production investigates how speakers produce grammatically well-formed utterances that communicate an intended message. the successful production of an utterance entails that, during grammatical encoding, the speaker must match not only lexical and morphological items to conceptual information from the message she intends to convey, but also that each of these items are compatibly integrated in a conventional syntactic structure that can then be phonologically encoded. as with other areas of psycholinguistics, research into sentence production is informed by data provided first through observation of the linguistic phenomenon in question and then experimental investigation and modeling of the phenomena under examination. as it is assumed that the psychological mechanisms involved in speech production are the same for all normal speakers, theories of these mechanisms must account for data observed in a wide range of typologically variant languages. the purpose of this i would like to thank alexandra aikhenvald, bhuvana narasimhan, susan brown, les sikos, steve duman, michael thomas, and two anonymous reviewers for feedback and discussion of the ideas presented here. all errors are of course my own. 1 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 2 paper is to present a general overview of current psycholinguistic approaches to one particular linguistic phenomenon, agreement, while drawing attention to data in a language unlike those previously considered in psycholinguistic studies of agreement, shipibo-konibo. it will be argued that the data observed in shipibokonibo, based on the work of pilar valenzuela (2002), are of interest not only because they exhibit features not seen in languages previously examined in agreement studies thus far, but also because they provide suitable stimuli that may be used in well-established experimental paradigms used to investigate agreement in speech production. so what is “agreement”? in theoretical linguistics, agreement is typically understood as an asymmetric syntactic relationship in which the form of one element (the “target”) in a sentence corresponds to the form of another (the “controller”) (corbett 2006). typical examples include number marking on verbs to correspond with the number of the subject, as in the english examples (1-2). (1) the cat (sg) plays (sg) (2) the cats (pl) play (pl) other features often considered as reflecting agreement in subject-verb relations include person and gender (but see corbett 2006:133-5 for discussion). within psycholinguistic studies of agreement production, one main question concerns the extent to which agreement morphology is influenced by information in the conceptual representation of the message rather than being strictly the result of syntactic procedures as defined by a language’s grammar. in other words, do targets (verbs) “look into” the conceptual message to access the notional values of agreement features, such as whether or not the referent is conceived of as singular or plural with regard to the number feature, or do they simply copy the grammatical values from the corresponding controlling elements (subjects) in the sentence, that is, whether or not the lexical item referring to the controlling element is specified as singular or plural1? thus, data of interest to studies of agreement production often include examples in which there is a mismatch between the notional value of the feature and the grammatical value of the feature. example of such mismatch with regard to the number feature include 1 in this example, ‘singular’ and ‘plural’ refer to values for the feature ‘number.’ however, the same distinction between ‘notional’ and ‘grammatical’ values is relevant for other agreement features, such as gender. in languages that mark gender, some referents (such as humans and other animate entities) may have notional gender values (e.g., female-feminine, male-masculine), while other referents may only have grammatical gender. this paper will focus on the number feature involved in subject-verb agreement, as that is the feature relevant to the shipibo-konibo data presented. 2 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 3 the english noun scissors, in which the referent is notionally singular but grammatically plural, or family, which is grammatically singular but may, in some dialects, have a notionally plural value (being conceived of as a set of indivual members). cases in which agreement morphology reflect the grammatical number of the controlling referent (the scissors are) are taken to be evidence for agreement production being governed by syntactic processes. on the other hand, when agreement morphology reflects the notional value of the controlling referent (the family are, in some dialects), we have evidence that conceptual information is relevant to the agreement production process. on one side of the debate are production models that describe agreement as being driven primarily by syntactic procedures. one such model is the marking and morphing model (eberhard, cutting & bock 2005). the marking and morphing model assumes a grammatical encoding process that includes roughly two components: functional assembly, during which lexical entries are accessed and matched to grammatical functions as marked by the conceptual message, and structural integration, at which point agreement morphology is added to the lexical forms that have been accessed, and those forms are integrated into the appropriate constituent structure. agreement processes operate under syntactic guidance with respect to hierarchical representations of sentence structure, where features are transmitted or copied from the controller to the target. during subject-verb agreement production the agreement target (verb) has no access to the conceptual representation of the controlling referent, but only to the grammatical value of the features as marked on the lexical form (the subject noun phrase, after it is encoded lexically). on the other end of the spectrum are constraint-based models such as the maximal input model (also referred to as the unification model; vigliocco, butterworth & semenza 1995; vigliocco, butterworth & garrett 1996; vigliocco & franck 2001; vigliocco & hartsuiker 2002). such models claim that agreement features marked on targets are derived not solely from the syntactic controller, but also from information in the conceptual representation of the message. targets have direct access to the notional value of the referent—for example, in the case of the number feature, whether or not the referent is conceived of as ‘singular’ or ‘plural.’ controllers and targets are marked separately for features and are then unified during structural assembly. during this unification process, agreement features are checked for compatibility.2 2 based on her work with franck (franck, vigliocco, antón-méndez, collina & frauenfelder 2008), one may assume that vigliocco now rejects the unification model. this rejection is based partially on the fact that the model has not fully accounted for morphophonological effects on agreement, but primarily on the claim that a conceptually-driven account of agreement “is 3 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 4 both models are based in part on observational data. the marking and morphing model accounts for the observation that subject noun phrases differing in notional number but having the same grammatical (e.g., the label on the bottles vs. the road to the lakes) both display grammatical agreement in english (bock & miller 1991). observational data that motivate the maximal input/unification model include agreement features marked on verbs in null-subject languages and conceptual effects on verb agreement features (vigliocco, butterworth & semenza 1995:188-189; vigliocco, butterworth & garrett 1996:264-266). beyond observational data, each model has been supported by a variety of experimental data, almost all of which is based on eliciting a type of speech error referred to as attraction (bock & miller 1991). attraction occurs when agreement features on a target erroneously match those on a referent other than the controller, as in the cost of the improvements have not yet been estimated, where have agrees with the plural improvements rather than the singular cost. yet the data investigated in studies underlying these models hardly cover all agreement phenomena. as eberhard, cutting and bock (2005:553) themselves note in presenting the marking and morphing model, “[n]o other models have yet been developed to address in any detail the range of findings generated in the literature on grammatical agreement, so there is ample room for improvement”. evidence for the models discussed are based primarily on results from empirical studies in english, although other languages have, to various degrees, been investigated (english: bock, eberhard & cutting 2004, bock, butterfield, cutler, cutting, eberhard & humphreys 2006; eberhard, cutting & bock 2005; french: franck, vigliocco, antón-méndez, collina & frauenfelder, 2008; german: berg, 1998; russian: lorimor, bock, zalkind, sheyman & beard 2008; hebrew: deutsch & dank 2008). agreement morphology in languages that are more typologically variant has not yet been examined. moreover, there is still much variation in the type of syntactic structures to be examined; while there is a well established literature on agreement in tag questions and subject-verb agreement in non-embedded clauses, (bock, nicol & cutting 1999; vigliocco, butterworth & semenza 1995; vigliocco, butterworth & garrett 1996) psycholinguistic research in agreement production has just begun to consider a wider variety of structures (see franck, frauenfelder & rizzi 2007). incompatible with most modern linguistic accounts of agreement which, in order to account for a number of syntactic phenomena, assume a fundamental difference between the way features are specified on the noun and on the verb or adjective” (franck et al. 2008:355). this critique ignores constraint-based accounts of syntax (pollard & sag 1994; wechsler & zlatic 2003). because no psycholinguistic model of agreement has yet explained the full range of agreement phenomena observed across languages, and because there are indeed modern syntactic theories that are compatible with a unification-based model of agreement, i take the maximal input/unification model to still be relevant. 4 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 5 how to address these gaps in the current literature? further development of psycholinguistic models to handle a wider range of agreement phenomena seen in language production can be based on two possibilities: one, considering languages that are more typologically variant, and two, following the current trend and continuing to examine structures that have not been previously examined with respect to agreement in the context of the model. this paper presents data that address both possibilities by considering the morphology of a particular linguistic structure, relative clauses, from data in a language previously uninvestigated in psycholinguistic studies—shipibo-konibo. it will be argued here that the examination of agreement morphology as well as other morphologically-marked relations of compatibility in relative clauses in shipibokonibo, a panoan language spoken in the peruvian amazon, challenges the architecture and underlying assumptions of current psycholinguistic models of agreement in speech production. shipibo-konibo is a morphologically rich language with variable plural marking on verbs, adverbial transitivity agreement, and case-marked arguments, among many other morphological features. within relative clauses, shipibokonibo demonstrates multiple positional types (pre-nominal, post-nominal, internally-headed) as well as various relativization strategies (including gaps and anaphoric pronouns). thus, the abundance of overt morphology and variation in shipibo-konibo relative clauses as compared to languages such as english offers an opportunity for empirical researchers to examine a broader range of notional and grammatical agreement morphology. such variance would allow for researchers to investigate the psycholinguistic processes of agreement production while varying more parameters, including word order, clause boundaries, and optional expression of morphemes. the remainder of this paper will be organized as follows: section 2 will present a brief typological sketch of data in shipibo-konibo, based on the work of pilar valenzula (2002), relevant to the discussion of the psycholinguistic mechanisms behind the production of agreement. section 3 provides a basic overview of the two psycholinguistic models of agreement in sentence production compared here, as well as how they differ with respect to the role of conceptual (notional) and syntactic (grammatical) information in the production of subjectverb agreement. section 4 will discuss shipibo-konibo relative clause features that show promise as data for further investigation in researching agreement in speech production, and section 5 briefly concludes. 2. relative clauses in shipibo-konibo 5 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 6 this brief typological sketch of shipibo-konibo, a panoan language with approximately 26,000-30,000 speakers inhabiting the peruvian amazon along the ucayali river and its tributaries, is based on the work of pilar valenzuela (2002). i will follow valenzuela’s operational definition of relative clauses as “all expressions in which an optional clause containing a verb form adds information about a single head nominal, even if the latter remains unexpressed,” (valenzuela 2002:6). three general characteristics of shipibo-konibo that will be relevant to the discussion of models of agreement production will be presented here: features of the morphological system, including the behavior of s and a arguments and number marking on verbs; flexible word order within both main clauses and relative clauses; and a range of relativization strategies. 2.1. morphological features relevant to subject-verb agreement shipibo-konibo has an ergative/absolutive phrasal-suffix case marking system. as there are no cross-referencing subject and object pronouns on verbs or auxiliaries in shipibo-konibo, and word order is relatively flexible, these casemarking suffixes are helpful in marking arguments of the verb. case marking is realized on main-clause arguments, which may be modified by relative clauses, as well as arguments within the relative clause. in the case of modified main-clause arguments, the case marking appears at the end of the noun phrase, attached to the relative clause modifying the argument. examples are shown in (3-6).3 in (3), ainbo “woman” is shown in the absolutive form, being the s argument. in (4), the ergative marker tonin is attached at the end of the relative clause modifying the a argument, ainbo, rather than at the end of ainbo. regarding case marking for arguments within the relative clause, (5-6) demonstrate the use of the ergative form e-n-ra “i” for the a argument in a single clause “i met a woman last year,” in (5), and the same use of the ergative marker when that clause is then embedded as a relative clause modifying ainbo “woman” in (6): (3) ainbo-ra kako-nko-ni-a-x nokó-ke woman:abs-ev kako-loc-lig-abl-i meet:dtrnz-cmpl ‘the/a woman arrived from kako.’ (valenzuela 2002:12) (4) ainbo [kako-nko-ni-a-x nokot-a]-tonin-ra rao woman kako-loc-lig-abl-i arrive-pp2-erg-ev plant.medicine:abs kobin-ak-[a]i boil-do.t-inc 3 a list of abbreviations and conventions (valenzuela 2002) are provided in the appendix 6 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 7 ‘the woman who arrived from kako is boiling the plant medicine.’ (valenzuela 2002:12) (5) e-n-ra ainbo onan-yantan-ke makáyain-xon 1-erg-ev woman:abs know-pst3-cmpl makaya:loc-t ‘i met the/a woman in makaya last year.’ (valenzuela 2002:13) (6) ainbo [makáyain-xon e-n onan-yantaan-a]-ra ne-no nokó-ke woman makaya:loc-t 1-erg know-pst3-pp2:abs-ev prox-loc meet:dtrnz-cmpl ‘the woman i met in makaya last year arrived here.’ (valenzuela 2002:13) although shipibo-konibo displays ergative/absolutive morphology, the language often treats subjects of intransitive verbs (s arguments) similarly to subjects of transitive verbs (a arguments). one example of such categorization is seen in plural agreement marking. plurality is coded through a verbal suffix if it is not indicated on the s/a argument (valenzuela 2002:1). the data presented here suggest that plural marking on the verb is not obligatory when the s/a argument is marked as in (7), as there is no plural marking on the verb keyo-ai “finish”, while there is plural marking on the ergative (a) argument joni-baon-ra “people”. it is required on the verb when the plural subject is omitted, as in the headless relative clause shown in (8); note the plural –kanon meni-kati-kan-ai “give”, and the absence of an ergative argument “they”. it is also required when the s argument is unmarked, as in (9), where plural is unmarked on bake “child” but is shown on be-kan-a “come”. the data also suggest that nothing prevents plural marking on the verb when it is marked on the subject nominal as well, as seen in (10), where both the a argument, shipi-baon-ra “the shipibo” and the verb pi[y]ama-kan-ai “eat” display plural morphology. (7) [jatik-xon-bi sepa-[a]i] joni-baon-ra jatí-tian ishton altogether-t-em slash-pp1 person-pl:erg-ev all-temp quickly keyo-ai finish-inc ‘men who slash a chacra altogether always finish quickly.’(valenzuela 2002:13) (8) [jawerato-n-ki yokat-ai] ja meni-kati-kan-ai. which-erg-int ask-pp1:abs 3:abs give-pst4-pl-inc ‘they gave her (her daughter) to whoever asked for (her).’ (valenzuela 2002:58) 7 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 8 (9) jain-ribi-ronki be-kan-a iki… oa bake there-rep-hsy come.pl-pl-pp2 aux dist child [moa xontako-ai]. already (become) young.woman-pp1:abs ‘there also came…those girls (who were) already turning into young women.’ (valenzuela 2002:55) (10) shipi-baon-ra kapé pi-[y]ama-kan-ai shipibo-pl:erg-ev alligator:abs eat-neg-pl-inc ‘the shipibo don’t eat alligator.’ (valenzuela 2002:9) it is the case for relative clauses as well as main clauses that when the plural s/a argument is not overtly expressed, plural marking on the verb is obligatory, as seen in (11). within the relative clause, the a argument “they” is not expressed. the verb, ta-nini-nan-yama-bain-wan-kan-a appropriately displays the plural morpheme –kan-. (11) nokon koka r-iki [jawen ochíti pos1 maternal.uncle:abs ev-cop pos3 dog:abs ta-nini-nan-yama-bain-wankan-a] joni foot-pull-mal-neg-and2-pst1-pl1-pp2 person ‘the man whose dog they did not pull by the foot to his detriment while passing earlier today is my maternal uncle.’ (valenzuela 2002:10) 2.2. word order the basic constituent order of shipibo-konibo is aov/sv, although word order within the main clause can be flexible to include non-verb-final orders. word order within the relative clause, however, is strictly verb-final, although a and o arguments may be switched (valenzuela 2002:15-17). within noun phrases, nouns and modifying elements including adjectives, quantifiers, numerals and relative clauses display flexibility in word order. (valenzuela 2002:1). shipibo-konibo displays an interesting variety of positional types of relative clauses. pre-nominal, post-nominal, and internally-headed relative clauses are found in the language, as well as relative clauses which separate a determiner and the head noun. discontinuous relative clauses are also present, in which the head and the relative clause are separated by a modifying expression. examples of pre-nominal, post-nominal and internally-headed relative clause types are presented in (12-14), where the head nominal jono “peccary” is shown following, preceding, and residing within the relative clause papa-n reteibat-a “(that) father killed”: 8 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 9 (12) [papa-n rete-ibat-a] jono-ra moa no-n keyo-ke father-erg kill-pst2-pp2 peccary:abs-ev already 1p-erg finish-cmpl ‘we already finished the collared-peccary father killed yesterday.’ (valenzuela 2002:18) (13) jono [papa-n rete-ibat-a]-ra moa no-n keyo-ke peccary:abs father-erg kill-pst2-pp2:abs-ev already 1p-erg finish-cmpl ‘we already finished the collared-peccary father killed yesterday.’ (valenzuela 2002:19) (14) [papa-n jono rete-ibat-a]-ra moa no-n keyo-ke father-erg peccary:abs kill-pst2-pp2:abs-ev already 1p-erg finishcmpl ‘we already finished the collared-peccary father killed yesterday.’ (valenzuela 2002:19) 2.3. relativization strategies in addition to the variety of positional types of relative clauses, shipibokonibo has several relativization strategies. these include a gap strategy, in which the relativized element corresponding to the head nominal is omitted from the relative clause, and an anaphoric pronoun strategy, where the relativized element corresponding to the head nominal is expressed in the relative clause by an anaphoric pronoun ja, a third-person singular pronoun unmarked for gender (valenzuela 2002:51). overall, while relative clauses in shipibo-konibo show some nominalization behaviors (some constraints on word order with the relative clause, and other nominalization features that will not be relevant to the analysis presented here), it is important to note that they exhibit hallmarks of main declarative clauses. like declarative clauses, relative clauses usually keep their full array of case-marked arguments and full adverbials with transitivity marking. most important for this study is the fact that relative clauses, like main clauses, show some flexibility regarding word order as well as their position relative to the modified head noun. thus, agreement morphology linking arguments and verbs is not dependent upon linear order or overt expression of arguments. the next section continues with an examination of the psycholinguistic models that will may examined with respect to such agreement morphology. 9 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 10 3. psycholinguistic models of agreement in sentence production 3.1. the marking and morphing model in the marking and morphing model (bock, eberhard, cutting, meyer & schriefers 2001; eberhard, cutting & bock 2005) subject-verb agreement in speech production is composed of two distinct stages during grammatical encoding. the first stage, marking, is a mapping of agreement features (such as person, number, and gender) from a conceptual representation to grammatical representations early in the grammatical encoding process. early in the grammatical encoding process, marking assures that a subject noun phrase is specified for agreement features in accordance with the conceptual representation. in certain cases, the features marked on the subject np may not be realized in the lexical specification. for example, in the sentence, “the sheep are grazing,” sheep has no morphological plural marker. yet the subject np the sheep is marked as notionally plural, as demonstrated by the verb morphology (i.e., are, rather than is). the realization of number (and, presumably, person and gender) on subject nps is then a joint product of the notional number retrieved from the conceptual representation and the grammatical number specified by the lexical representations used to build the noun phrase. a computational version of the theoretical marking and morphing model explains how both notional number (from marking) and grammatical number (from lexical specifications) contribute to the final value of number for subject noun phrases (eberhard, cutting & bock 2005). while lexical specifications of local nouns inside the subject noun phrase (i.e., books in “the editor of the history books,”) are calculated into the final subject np number value (and, in some cases, can override the head noun, leading to errors in agreement), notional number of local nouns is not a factor. moreover, the lexical specification of nouns embedded in clausal modifiers inside the noun phrase (i.e., books in “the editor who rejected the books,”) are less likely to affect the number value of the subject np than nouns in the same clause as the head noun of the subject np (bock & cutting 1992). in the second stage, morphing, subject noun phrases control the agreement marking on the target verb by copying person-number-gender features, as determined by a combination of the notional marking process and lexical specification, onto the verb phrase. this occurs later in the grammatical encoding process, at the point when agreement-relevant features marked on grammatical representations (i.e., ‘subject’ marked as ‘plural’) are reconciled to those features specified in the lexicon (either a plural morpheme ‘-s’ or the appropriate lexical item, as in ‘sheep’). those morphological forms are retrieved and integrated into the constituent structure in the subject noun phrase position. the realization of agreement morphology on the verb, however, is constrained by syntactic processes. verbs inherit person-number-gender features 10 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 11 from the subject noun phrase; they cannot directly access notional number from the conceptual message. the verb ‘morphs’ to take on the correct form in accordance with the number value copied from the subject noun phrase. thus, the marking and morphing model treats agreement as primarily driven by syntactic procedures in which agreement features are copied from the subject np to the verb during grammatical encoding. 3.2. the maximal input/unification model a second model for the production of agreement, the maximal input model (vigliocco, butterworth & semenza 1995; vigliocco, butterworth & garrett 1996; vigliocco & hartsuiker 2002), claims that the production of agreement morphology is semantically driven. unlike the marking and morphing model, in which only the grammatical representation of the subject np receives person-number-gender information from the message’s conceptual representation, in the maximal input model, relevant conceptual features are retrieved by both the subject np and the verb. in this approach, features are not copied from a controller to a target. rather, each element (in this case, the subject and the verb) individually retrieves information from the conceptual representation. agreement is thus a relation in which two elements supply partial information about a single linguistic form. unlike the marking and morphing model, the maximal input model does not assume directionality (i.e., a controller-target relationship) in agreement, although if one element involved in the agreement relationship carries more information than another, as in english, agreement may appear directional. once features have been retrieved from the conceptual representation by the head of the subject np and the verb, they are passed up to the highest projections in the syntactic structure (the subject np node and the vp node) and the structures then undergo a unification process. this occurs during the grammatical encoding stage when constituents are combined in a structural representation but before word order is determined. subject-verb agreement is the result of unification of the subject np and the vp at the s node (vigliocco, butterworth & garrett 1996). the final realization of agreement features is a combination of the information provided by both subject np and verb, and unification may be considered to be a sort of “feature checking” procedure that ensures that the features of each element are compatible (franck, vigliocco & nicol 2002:376). 4. shipibo-konibo relative clauses and psycholinguistic models of subjectverb agreement 11 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 12 as presented in section 2, above, the typological features of relative clauses in shipibo-konibo offer new potential for assessing psycholinguistic models of agreement in speech production. these features will be discussed here with respect to the models, looking at both observational data (attested utterances that it is assumed the models strive to account for), and data of experimental interest (potential stimuli that may be constructed in shipibo-konibo for the purposes of experimental investigation of the models.) 4.1. observational data relative clauses in shipibo-konibo provide an opportunity to study subject-verb agreement production that languages previously investigated with respect to agreement do not provide. unlike previously studied languages, shipibo-konibo displays a wide range of flexibility not only with regard to word order in both main and, to some extent, relative clauses, but also in the positional types of relative clauses allowed. as discussed above, plural morphology may or may not appear on the verb when plural subject nps (s/a arguments) are overtly expressed, but when the plural subject np is not expressed, plural marking on the verb is obligatory. the first issue, then, concerns the case of non-overt subjects. while this particular issue is not unique to shipibo-konibo and has been discussed previously in both psycholinguistic and theoretical linguistic studies of agreement, the data presented here allow new ways to address the issue. the relative clause verb in (15) shows the plural suffix –kanwhile the a argument “they” is not expressed overtly. (15) jain iki pionis bepon [ja-n rao-n-kati-kan-ai] there cop pionis resina 3-inst plant.medicine-trnz-pst4-pl-pp1:abs ‘there is the resina pionis with which they cured the girls.’ (valenzuela 2002:23) it has been suggested in previous literature that the phenomenon of agreement with null-subjects is problematic for hierarchically-based syntactic models of agreement, such as the marking and morphing model or the feature selection and copy model (which will not be examined here) (franck, vigliocco, antón-méndez, collina & frauenfelder 2008), as no subject exists in the utterance from which to copy agreement features. a unification-based approach such as that of the maximal input model, however, is more easily able to account for the appearance of agreement morphology on the target (the verb) in absence of the controller (the subject) because agreement does not depend upon the surface expression of the controller; the target may receive feature information directly from the conceptual message. 12 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 13 a somewhat different picture is painted by relative clause utterances in which the unexpressed subject corresponds to the modified head nominal, as in (16): (16) [jatik-ax-bi teet-ai] joni-baon-ra jatí-tian ishton keyo-ai altogether-i-em work-pp1 person-pl:erg-ev all-temp quickly finish-inc ‘people who work altogether always finish quickly.’ (valenzuela 2002:13) although the subject joni-baon “people” is not expressed in the relative clause, the verb teet-ai “work” does not display plural morphology. there are two possible explanations for this. first, it may be the case that shipibo-konibo only requires the plural-marked head nominal corresponding to the subject of the clause to be expressed in the main clause of the utterance in order to omit plural verbal morphology. second, as the relative clause displays a “gap” strategy of relativization, one might hypothesize that the syntactic structure of the relative clause contains a trace—a null argument that carries all of the features of the subject, even though it is not overtly expressed. the first hypothesis may very well be in line with some version of the constraint-based maximal input model but would be problematic for syntactically driven models. without a subject carrying features to control the form of the verb, there is no way to predict whether the verb should or should not exhibit the plural morpheme. the constraint-based model shows more promise. the lack of plural morphology on the verb in the relative clause can presumably be explained in maximal input model because of two features: one, the unification procedure that occurs during constituent assembly is a checking procedure to make sure that features expressed on the head nominal coindexed with the relative clause subject and relative clause verb are compatible; and, two, it is assumed that all that is necessary for such compatibility is for the plural marking to be expressed by at least either the head nomial or the relative clause verb. in this way, the maximal input model can explain examples such as (16) as well as cases in which both subject and verb show plural marking, and those in which there is plural marking on the verb in the absence of marking on the subject. this would require, however, that the gap in the relative clause (the missing subject) be coindexed with the head nominal. while unification based theories of syntax posit such representations where a missing element can be coindexed with other elements in the utterance—without positing a “trace” (see sag, wasow & bender 2003, chp. 14)—the maximal input model has not been fully developed to explain how such representations would be processed on-line during speech production. the second hypothesis is more compatible with syntactically-driven models of speech production. if one assumes an underlying representation of the relative clause contains a subject argument coindexed with a trace containing its 13 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 14 features, the form of the verb can be explained as a result of those features. evidence for traces controlling target features has been demonstrated in prior experimental research (franck, frauenfelder & rizzi 2007). the difference in shipibo-konibo, however, is that the plural features on the subject cause an omission of features being expressed on the verb rather than a copying of features on to it. this would not be terribly problematic—it would still be a “systematic covariance” of one element dependent upon another (steele 1978:610, cited in corbett 2006)—if it were not for the fact that shipibo konibo shows variability in the marking of subject-verb agreement. example (17) provides an example of a relative clause in which the omitted subject corresponds to the marked-plural head nominal, but the verb also displays plural marking: (17) jain ik-á iki oa [no-a shiro bewakan ninká-ma-ai-bo] there do.i-pp2 aux dist 1p-abs shiro song:obl hear-caus-pp1-pl ainbo-bo. woman-pl:abs ‘there stood the women who provoked us with their shiro songs.’ (valenzuela 2002:28) finally, shipibo-konibo relative clauses are of interest in assessing psycholingistic models of agreement in speech production in that they present data for which the models cannot account. one such example concerns number marking of resumptive pronouns within relative clauses. as shown in (18), the resumptive pronoun ja-n (3-erg) that corresponds to the head nominal joni-bo “men”, is singular, despite its reference to a plural antecedent. (18) [ja-n jato bi-ai] joni-bo ik-á iki tampóra-ya. 3-erg 3p:abs get-pp1 person-pl:abs be-pp2 aux drum-prop ‘those men who welcomed them had drums.’ (valenzuela 2002:59) this mismatch in number marking between pronoun and antecedent cannot be explained as a difference in notional and grammatical number of the target, nor can it be described as attraction. previous research on pronounantecedent agreement has have treated the phenomenon as being more sensitive by conceptual information than by grammatical features; this is not the case here. anaphoric pronouns in shipibo-konibo relative clauses are generally rejected by speakers (valenzuela 2002:58). both syntactically-driven and constraint-based models of agreement production have yet to explain how agreement features might be blocked from targets in certain clauses that would otherwise express those features. that such data have not yet been explained by the models is not in 14 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 15 and of itself surprising; it is, however, interesting in that tokens such as these suggest directions for future work in modeling agreement in speech production. 4.2. data of experimental interest the variation of relative clause types and the flexible word order in shipibo-konibo allow for the creation of possible stimuli that can test and develop the agreement mechanisms proposed by psycholinguistic models using well-established experimental paradigms. for example, consider the verb and pronouns within the pre-nominal relative clause with respect to the head nominal in (19): (19) tita-r keyot-ai, ja-tian no-a bane-ti ka-[a]i mother;abs-ev finish-inc that-temp 1p-abs stay-inf go-inc [ja-n no-a axe-a-a] jawéki-bo-ya. [3-erg 1p-abs get.used.to-caus-pp2] thing-pl-prop ‘our mother dies and then we stay with the things she has taught us.’ (valenzuela 2002:19) the ergative pronoun ja-n, in agreement with its antecedent, tita-r “mother”, has no plural marking. likewise, the verb also lacks plural marking. the plural head nominal jawéki-bo-ya “things” follows the prenominal relative clause. but shipibo-konibo allows for not only pronominal but also post-nominal and internally-headed relative clauses. the inventory of relative clause types of shipibo-konibo presumably would allow the head nominal jawéki-bo-ya “things” to appear inside the relative clause, between the subject ja-n ‘she’ and the verb axe-a-a “taught.” would it be possible to elicit attraction errors in shipibokonibo internally-headed relative clauses by placing a head nominal between the subject and verb of a relative clause, resulting in a plural marking on the verb (in bold), as hypothesized in (20)? (20) tita-r keyot-ai, ja-tian no-a bane-ti ka-[a]i mother;abs-ev finish-inc that-temp 1p-abs stay-inf go-inc [ja-n no-a jawéki-bo-ya axe-a-a (-kan-?)] [3-erg 1p-abs thing-pl-prop get.used.to-caus-pp2 (-pl-?)] ‘our mother dies and then we stay with the things she has taught us.’ (adaptation my own) 15 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 16 by allowing variations in form for the same semantic content as seen in (20), shipibo-konibo provides a rare opportunity to tease apart syntactic and semantic factors in the production of agreement.4 there should be no difference between the semantics of (19) and (20) above; therefore, any attraction effects seen would have to be the result of syntactic processes and not conceptual features. the possibility of constructing experimental stimuli to elicit attraction effects is further suggested by the existence of discontinuous relative clauses such as (21), in which the head nominal is separated by a clause with an attributive function: (21) ja-káti-ai [yotokoni pi-á] kikin xeta wiso-bi-ribi exist-pst4-inc yokotoni:abs eat-pp2 extremely tooth black-em-also ik-í joni-bo do.i-sssi person-pl:abs ‘there were people who ate yotokoni and whose teeth were extremely (shiny) black.’ (valenzuela 2002:30) in (21), the a argument corresponding to the head nominal joni-bo ‘people’ is not expressed within the relative clause. however, the relative clause verb pi-á ‘eat’ has no plural marking despite the separation of relative clause from head nominal. just as (18) above, (21) demonstrates that it is not the case that such marking is required on a relative clause verb when a subject argument is missing in the relative clause, as long as that missing subject corresponds to the head nominal expressed in the main clause. what is worth noting here is that the variety of elements that are allowed to appear between relative clause and head nominal provide an opportunity for designing stimuli that would be useful in examining attraction phenomena in the production of agreement. experiments examining attraction in english number agreement have found that attractors specified for number are more likely to affect agreement than attractors that have a default number. in english, singular nouns, which are unmarked, are less likely to cause an attraction effect than nouns marked for plurality (bock et al. 2001). number in shipibo-konibo appears to be similarly marked in that the plural is specified while the singular form is the default. the agreement pattern, however, defies the canonical definition of agreement (corbett 2006). it is not the presence of the number feature on the noun that requires a number marking on the verb; rather it is the absence of the feature that triggers number agreement. the presence of the plural marker on the noun actually seems 4 whether or not variations in the form of relative clauses in shipibo-konibo are due to discourse factors has not yet been investigated (valenzuela 2002:51). 16 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 17 to make plural marking on the verb optional. how, then, might it be possible to test attraction effects in this language? one possibility would be to create a variation of stimuli based on those used in previous experiments, in which singular subject noun phrases contain plural attractors, but to have the plurality unmarked on the attractor such that it might trigger the obligatory plural marking on the verb at a higher rate than singular attractors. consider (22), an assumed adaptation of (21) above, in which the head nominal “people” would be placed before the relative clause, and xeta (assumed to be notionally plural “teeth” in this context) were to also remain unmarked for plurality5: (22) ja-káti-ai joni-bo kikin xeta wiso-bi-ribi… exist-pst4-inc person-pl:abs extremely tooth black-em-also … ‘there were people whose teeth were extremely (shiny) black and who…’ (adaptation my own) now consider a token in which xeta “teeth” were to be replaced with a notionally singular item, perhaps the word for “canoe,” nonti: (23) ja-káti-ai joni-bo kikin nonti wiso-bi-ribi… exist-pst4-inc person-pl:abs extremely canoe black-em-also … ‘there were people whose canoe was extremely (shiny) black and who…’ (adaptation my own) as with previous experiments investigating attraction in subject-verb agreement, the point of interest is whether or not speakers prompted with sentence fragments given above finish the sentence with a plural form of the verb. because the subject joni-bo “people” is marked as plural, a verb either marked or unmarked for plurality would be acceptable. because unmarked plural subject nouns require verbs marked for plurality, however, if there is a higher rate of verbs produced with plural marking in the presence of a notionally plural attractor (xeta “teeth”) than singular attractors (nonti “canoe”), we may conclude that shipibo-konibo provides evidence in support of a constraint-based psycholinguistic model of agreement production. a second possibility for experimental investigation lies in the variation of positional types of relative clauses in shipibo-konibo. as discussed in section 2, shipibo-konibo displays pre-nominal, post-nominal and internally-headed 5 the examples adapted from valenzuela (2002) are intended for illustration purposes and may not be grammatically felicitous. naturally, any stimuli created for experimental investigation would require the review of a native speaker consultant. 17 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 18 relative clauses. consider possible variations of example (19), repeated here as (24), as containing a post-nominal (25) and internally-headed relative clause (26) with the head nominal as unmarked for plurality and the verb omitted. (24) tita-r keyot-ai, ja-tian no-a bane-ti ka-[a]i mother;abs-ev finish-inc that-temp 1p-abs stay-inf go-inc [ja-n no-a axe-a-a] jawéki-bo-ya. [3-erg 1p-abs get.used.to-caus-pp2] thing-pl-prop ‘our mother dies and then we stay with the things she has taught us.’ (valenzuela 2002:19) (25) tita-r keyot-ai, ja-tian no-a bane-ti ka-[a]i mother;abs-ev finish-inc that-temp p-abs stay-inf go-inc [ja-n no-a jawéki-ya … [3-erg 1p-abs thing-prop … ‘our mother dies and then we stay with the things she ____ us.’ (adaptation my own) (26) tita-r keyot-ai, ja-tian no-a bane-ti ka-[a]i mother;abs-ev finish-inc that-temp 1p-abs stay-inf go-inc jawéki-ya [ja-n no-a … thing-prop [3-erg 1p-abs … ‘our mother dies and then we stay with the things she ____ us.’ (adaptation my own) both the post-nominal and internally-headed relative clauses allow the head nominal to appear before the verb, making it possible to examine any attraction effects it might have on the production of the verb. similar to the previous example, the influences of syntax and semantics in the production of agreement in speech might be examined by replacing the notionally plural head nominal attractor with a notionally singular attractor and examining the rates of production of plural marking on the elicited verb. these are just a couple of examples of the range of possible experimental stimuli that might be designed in shipibo-konibo to examine psycholinguistic processes of subject-verb agreement production. additional possibilities include examining attraction effects with respect to variations in linear and structural distance, examining the role of clause boundaries in attraction effects, and attraction effects on main verbs. 18 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses 19 5. conclusion even the cursory examination of typological features of relative clauses in shipibo-konibo presented here clearly demonstrates that this language is unique among those considered for psycholinguistic investigation. this reason alone should prove enough to prompt researchers to add it to the lists of languages under psycholinguistic investigation up to now. but even beyond the need to consider a wider range of language types in psycholinguistic research, the variation seen in shipibo-konibo with respect to word order, relative clause positional types and relativization strategies are absolutely compelling. these features provide a wealth of possibilities for investigating both syntactic and semantic influences on agreement in speech production. moreover, experimental paradigms such as elicitation techniques used in numerous previous studies promise to be suitable for investigating shipibo-konibo. 19 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 20 references berg, thomas. 1998. “the resolution of number conflicts in english and german agreement patterns.” linguistics 36(1): 41-70. bock, kathryn, & carol a. miller. 1991. “broken agreement.” cognitive psychology 23(1): 45-93. bock, kathryn, & j. cooper cutting. 1992. “regulating mental energy: performance units in language production.” journal of memory and language 31(1): 99-127. bock, kathryn, janet nicol, & j. cooper cutting. 1999. “the ties that bind: creating number agreement in speech.” journal of memory and language 40(3): 330-346. bock, kathryn, kathleen m. eberhard, j. cooper cutting, antje s. meyer, & herbert schriefers. 2001. “some attractions of verb agreement.” cognitive psychology 43(2): 83-128. bock, kathryn, kathleen m. eberhard, & j. cooper cutting. 2004. “producing number agreement: how pronouns equal verbs.” journal of memory and language 51(2): 251-278. bock, kathryn, sally butterfield, anne cutler, j. cooper cutting, kathleen m. eberhard & karin r. humphreys. 2006. “number agreement in british and american english: disagreeing to agree collectively.” language 82(1): 64113. corbett, greville g. 2006. agreement. new york: cambridge u press. deutsch, avital, & maya dank. 2008. “conflicting cues and competition between notional and grammatical factors in producing number and gender agreement: evidence from hebrew.” journal of memory and language 60: 112-143. eberhard, kathleen m., j. cooper cutting, & kathryn bock. 2005. “making syntax of sense: number agreement in sentence production.” psychological review 112(3): 531-559. franck, julie, gabriella vigliocco, & janet nicol. 2002. “subject-verb agreement errors in french and english: the role of syntactic hierarchy.” language and cognitive processes 17(4): 371-404. franck, julie, ulrich h. frauenfelder, & luigi rizzi. 2007. “a syntactic analysis of interference in subject-verb agreement.” mit working papers in linguistics 53: 173-190. franck, julie, gabriella vigliocco, inés antón-méndez, simona collina, & ulrich h. frauenfelder. 2008. “the interplay of syntax and form in sentence production: a cross-linguistic study of form effects on agreement.” language and cognitive processes 23(3): 329-374. lorimor, heidi, kathryn bock, ekaterina zalkind, alina sheyman, & robert beard. 2008. “agreement and attraction in russian.” language and cognitive processes 23(6): 769-799. 20 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 subject-verb agreement in the production of shipibo-konibo relative clauses 21 sag, ivan a., thomas wasow, & emily m. bender. 2003. syntactic theory: a formal introduction (2nd ed.). stanford, calif: center for the study of language and information. steele, susan. 1978. “word order variation: a typological study.” in j. a. greenberg, c. a. ferguson & e. a. moravcsik (eds.) universals of human language iv: syntax, 585-623. stanford: stanford university press. valenzuela, pilar. 2002. relativization in shipibo-konibo: a typologicallyoriented study. münchen: lincom europa. vigliocco, gabriella, brian butterworth, & carlo semenza. 1995. “constructing subject-verb agreement in speech: the role of semantic and morphological factors.” journal of memory and language 34(2): 186-215. vigliocco, gabriella, brian butterworth, & merrill f. garrett. 1996. “subjectverb agreement in spanish and english: differences in the role of conceptual constraints.” cognition 61(3): 261-298. vigliocco, gabriella, & julie franck. 2001. “when sex affects syntax: contextual influences in sentence production.” journal of memory and language 45(3): 368-390. vigliocco, gabriella & robert j. hartsuiker. 2002. “the interplay of meaning, sound & syntax in language production.” psychological bulletin 128: 442472. 21 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 22 appendix: abbreviations and conventions (valenzuela 2002) 1 first person singular 2 second person singular 3 third person singular 1p first person plural 2p second person plural 3p third person plural a transitive subject function abl ablative abs absolutive advz adverbializer agtz agentivizer all allative assoc associative att attenuative aug augmentative aux auxiliary ben benefactive caus causative chez chezative cmpl completive aspect com comitative coni conjunction cop copula des desiderative dim diminutive dist distal distr distributive dtrnz detransitivizer dub dubitative em emphatic erg ergative ey direct evidential fds following event, different subject fssi following event, same-subject, intransitive matrix clause fsst following event, same-subject, transitive matrix clause fut future gen genitive hab habitual hsy hearsay hsy2 shorter hearsay i intransitive (subject orientation) imp imperative inc incompletive aspect inf infinitive infr inferential inst instrumental int interrogative intens intensifier intrst complement of interest lig ligature lim limitative loc locative mal malefactive mns means neg negative nmlz nominalizer nom nominative o object function obl oblique onom onomatopoeia pds previous event, different subject pel pejorative p/j prospective/jussive pl plural po>s/a previous event, dependent object is coreferential with matrix subject posl possessive first person singular pos3 possessive third person singular 22 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/2 doi: https://doi.org/10.25810/j4en-sm72 subject-verb agreement in the production of shipibo-konibo relative clauses 23 ppl incompletive participle pp2 completive participle prev preventive priv privative prog progressive prop proprietive pssi previous event, samesubject, intransitive matrix clause psst previous event, samesubject, transitive matrix clause pst1 earlier today past pst2 yesterday past pst3 several months/a few years ago past pst4 several years ago past rec reciprocal rem remote past s intransitive subject function sds simultaneous event, different subject siml similitive specl speculative sssi simultaneous event, same-subject, intransitive matrix clause ssst simultaneous event, same-subject, transitive matrix clause t transitive (subject orientation) temp temporal trnz transitivizer voc vocative 23 duffield: developing psycholinguistic models of subject-verb agreement in speech production published by cu scholar, 2010 colorado research in linguistics 6-2010 developing psycholinguistic models of subject-verb agreement in speech production: new data from shipibo-konibo relative clauses cecily jill duffield recommended citation microsoft word cril_duffield_revisions_submit.doc a corpus-based linguistic analysis of latin frequentative verbs a corpus-based linguistic analysis of latin frequentative verbs cover page footnote this paper was originally completed as part of the university of colorado department of linguistics preliminary examination doctoral requirement, submitted to the preliminary examination committee on october 1, 2015 and passed on december 3, 2015. i would like to thank dr. laura michaelis for helping develop and problematize the introduction of this paper. this working paper is available in colorado research in linguistics: https://scholar.colorado.edu/cril/vol24/iss1/1 https://scholar.colorado.edu/cril/vol24/iss1/1?utm_source=scholar.colorado.edu%2fcril%2fvol24%2fiss1%2f1&utm_medium=pdf&utm_campaign=pdfcoverpages a corpus-based linguistic analysis of latin frequentative verbs jared desjardins university of colorado boulder considering that latin frequentative verbs have transparent morphological structure (the supine stem of the base verb concatenated with the frequentative -itare suffix and regular first conjugation inflectional endings), one would assume that the meaning of each such verb is related in a predictable way to that of its corresponding base verb, and that all members of the frequentative class share semantic entailments. however, frequentative verbs resist a uniform semantic analysis and can mean something entirely unpredictable from the sum of their parts, and traditional and contemporary definitions of frequentatives ignore the degree of idiomaticity between frequentative form and function and are based on limited corpus data. this paper provides a synchronic, corpus-based linguistic analysis of frequentative verbs in comparison to their base forms from both derivational (source-oriented) and usage-based (product-oriented) perspectives in order to explore the nature of the latin frequentative. the paper concludes that a usage-based, product-oriented treatment of the data provides a more straightforward characterization of latin frequentative verbs, highlighting the interconnectedness of morphology, syntax, and semantics. keywords: morphology, syntax, semantics, corpus linguistics, latin 1. introduction frequentative (fv) verbs (from latin frequentare ‘to repeat often’) comprise a class of latin verbs classically characterized as denoting forcible or repeated action, as in the case of pulsare ‘beat repeatedly’, from pellere ‘beat’. a salient feature of the latin lexicon, fv forms survive in the etymologies of learned english borrowings (e.g., inhabitant, fluctuate, agitate) and as a shared inheritance of romance languages, in numerous reflexes of latin fv lexemes: italian cacciare ‘hunt’, from the fv form of latin capio ‘catch’, french raser ‘shave’, from the fv form of latin radere ‘scrape’, and spanish cantar, from the fv form of latin canere ‘sing’ (solodow 2010:151-153). the latin fv template is both highly productive and transparent, representing a straightforward instance of concatenation: fvs are formed by first-conjugation inflection of the supine (passive-participial) stem of the base verb (greenough et al. 2001:152). while fvs are formed from all three latin verb conjugations, the derived fvs are exclusively first conjugation (and exhibit first conjugation morphology), as illustrated by rogitare (derived from first conjugation rogare ‘ask’), habitare (derived from second conjugation habere ‘have’) and cursare 1 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 (derived from third conjugation currere). a survey of fv instances in the classical corpus (comprising latin works from the 100 bce to 100 ce period) shows that the fv pattern is widely attested across that particular latin lexicon. the pattern’s internal transparency and high type frequency suggests that its semantic effect is equally transparent – that the fv derivation modulates the base verb’s semantic representation in a predictable and uniform way. it is immediately apparent, however, that the traditional account of fv meaning is inadequate; while some fv predications do express repeated actions, or actions performed with unusual force, many do not. the fv habitare, for example, is derived from a state verb (habere ‘to have, hold’), and does not indicate repeated episodes of having. instead, it typically denotes ‘holding’ a piece of real estate, or inhabiting a house, as demonstrated in 1: (1) et cn. servilio praetori urbano negotium datum ut campani cives, ubi cuique ex senatus consulto liceret habitare, ibi habitarent, animadverteretque in eos qui alibi habitarent ‘additionally cnaeus servilius, the city praetor, was to see that the campanian citizens were living where the senate allowed them to live, and he was to punish those living elsewhere’ (liv. 28.46) example 1 is an account of cnaeus servilius’s duties as city praetor, an elected government position in which he is to oversee that the citizens of capua are living only where the senate decrees and punish those who reside elsewhere. livy’s use of habitarent is in the context of the senate’s orders to servilius and his responsibilities as city praetor and servilius’s obligations to manage capua’s residential laws. a traditional fv interpretation of habitarent in 1 is problematic, since a reading in which the capuan citizens repeatedly engage in states of ‘having’ is inconsistent with the context set up by servilius’s position as city praetor, and his orders to enforce land ownership and residential restrictions. instead a more idiomatic interpretation of residing or inhabiting a house for habitarent is preferred, since ‘residing in/inhabiting a house’ can be seen as a metaphorical extension of the base verb habeo’s basic meaning of ‘having’. considering fv verb lexemes tend to be highly polysemous, which follows from the productivity of the fv template in the language, and that the meaning of a fv verb is often unpredictable in relation to its ‘base’ verb lexeme’s semantics (since new meanings are known to replace the original fv sense) the traditional grammatical treatment raises questions concerning what it means to be a fv, and how that verb class can be characterized in latin. to augment the 2 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 repertoire of notions traditionally associated with fv meaning, modern authors including viti (2012) have characterized fvs as expressing “emphasis or expressivity” (p. 1). one can generalize this characterization to one that applies to those instances in which fvs express actions performed in unusual ways. one such instance is given in 2: (2) singula verba vellenti tamquam dictaret non diceret single words plucking as if dictate not speak ‘when vinicius was dragging out his words one by one, as if he were dictating, not speaking’ (sen. ep. 40) example 2 is a passage from one of seneca’s letters to lucilius, in which the author is discussing the proper style for a philosopher’s speech and communication. here seneca recommends that lucilius not concern himself with the criticism of those who care more about the volume of output than the manner in which it is conveyed and suggests lucilius speak as publius vinicius ‘the stammerer’ speaks. the passage offers a contrast between the canonical mode of speaking, denoted by diceret and a hyperarticulate mode of speaking expressed by dictaret. the repertoire of possible fv meanings can therefore be seen as including more figurative variants of both their base verb’s semantics as well as the semantics associated with the derived fv, lending themselves more readily to be used creatively and in novel contexts and environments (both in terms of the formal characteristics of their distribution(s), as well as their use in different semantic and pragmatic contexts). this is supported by the observation that fv verbs are numerous in comedy (viti 2012:1). in other words, viti (2012:2) contests the classical assumption by contrasting the a priori traditional definition of the repeated, intensive fv interpretation with what she claims to be the fv derivation’s primary function of emphasis and expressivity as a strategy to express imperfective (either progressive or habitual) aspect and backgrounded information (particularly when taken into consideration with proto-indo european’s historical aspectual system paradigm). however, even viti’s broadened definition fails to apply to the wide array of fv forms. for example, viti’s labels of ‘emphasis’ and ‘expressivity’ are highly subjective, limiting the explanatory and predictive power that would follow from an operationalized, synchronic analysis utilizing formalized linguistic features. furthermore, even viti’s proposed expanded repertoire of 3 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 fv meanings fails to include a majority of highly idiomatic fv interpretations, as in examples 3 and 4: (3) domin-us callist-um vendidit master-nom callistus-acc sell ‘the master sold callistus’ (sen. ep. 47) (4) quint-us frater [ … ] tusculan-um venditat ut … quintus-nom brother.nom [ … ] tusculan.property-acc sell so.that… ‘my brother quintus is now trying to sell his tusculan property, in order to purchase, if he can, the townhouse of pacilius’ (cic. att. 1.14) in 3, seneca is recalling how much more callistus’s former master lost in comparison to his gains in selling callistus. seneca ends by saying dominus callistum vendidit ‘the master sold callistus [but how much has calllistus made (his master) pay for!]’, a standard transitive use of the base verb vendidit ‘he sold’. in 4, cicero is telling atticus, a close friend, that his brother quintus is trying to sell his tusculan property. the emphatic, intensive action can be inferred from the fact that the selling action (venditat) introduces a purpose clause (headed by the conjunction ut): so that he (quintus) can purchase a townhouse he desires. however, both a traditional as well as an emphatic/expressive (cf. viti 2012) interpretation result in inadequate translations, since the use of venditare in 4 is not only an intensive form of ‘selling’, but a very specific form of ‘selling property’. given that fvs resist a uniform semantic analysis and that any fv can mean something entirely unpredictable from the sum of its parts, and traditional and contemporary (viti 2012) fv definitions ignore the degree of idiomaticity between fv form and function (and are based on limited corpus data), this paper therefore provides a synchronic, corpus-based linguistic analysis of fv verbs in comparison to their base forms from both derivational (source-oriented) and usage-based (product-oriented) perspectives (bybee 2001:126) in order to explore the nature of the latin fv ‘derivation’. to do so, i analyze 15 tokens of each fv and corresponding base verb for the top nine most frequent fv lexemes (totaling 135 fv and base verb tokens; 270 verb tokens total), in terms of their morphological, syntactic, and semantic properties. the remainder of this paper is organized as follows: in section 2 i outline the procedure and methodology in acquiring the latin fv data and corpus development, and the subsequent linguistic analysis of 4 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 each verb token. in section 3 i present and discuss the results of the linguistic analysis and provide a summary and concluding remarks in section 4. 2. methodology and analysis 2.1. corpus data the present analysis utilizes a latin data corpus (currently in development) in order to locate any and all possible fv verb forms. first, the entire collection of the latin perseus digital library (crane 2005) (provided under the creative commons sharealike 3.0 license) was downloaded as individual extensible markup language (.xml) files. since the latin text in each .xml file was encoded in a machine-readable format, a python program was developed to extract all latin text, which was then written to individual text (.txt) files. once the latin .txt files were generated, a subsequent program was created to clean, format, and normalize the latin text, addressing the following considerations: • each generated .txt file contained various xml and html (hypertext markup language) encodings. therefore, a function was implemented which removed all remaining xml or html encodings, while leaving any latin text unchanged. • due to the frequency of roman dates (e.g. prid. id. mart. = pridie idus martias = march 14), roman numerals (e.g. lxvii = 67), and roman name abbreviations2, a separate function identified each (and their various forms) in order to treat them as proper constituents. • another common occurrence in latin text concerns the orthographic representation of the phonemes /j/ and /w/. for example, the verb /ˈjakio/ ‘i throw’ might be written as iacio or jacio. similarly, the noun /ˈserwus/ ‘slave’ might be represented as servus or seruus. therefore, the data was normalized so that the latin phonemes /j/ and /w/ were uniformly represented by the same grapheme (⟨i⟩ and ⟨u⟩, respectively). • finally, the latin data was generally formatted, addressing concerns such as capitalization and punctuation, and a new .txt file corresponding to each latin author and particular text (e.g. caesar: de bello gallico ‘on the gallic war’) was generated, producing the entire raw text latin corpus. 2 it was common in latin literature to abbreviate one’s praenomen, or first name, as in c. (gaius) julius caesar and m. (marcus) tullius cicero. 5 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 since the present study considers fv verbs from the golden age of latin (approximately 70 bce – 18 ce specifically), the corpus was partitioned based on authors who were active during that time period. at the time of writing, the golden age latin subcorpus consists of 16 authors (from caesar to vergil) and 157,8603 tokens. once the latin golden age subcorpus was sufficiently prepared, a final program was written in order to: (1) locate any and all possible fv verb forms, (2) generate separate results files containing each fv token and textual context, and (3) return the token frequency of each fv verb. since the derived fv verb form is invariably first conjugation, regardless of the conjugation class of the base verb, certain morphological information was taken into consideration in the detection process. as the program iterated over each token, it ‘checked’ whether the token is a first conjugation verb, based on predictable first conjugation verbal morphology (specifically, the presence of the vowel a in the verb stem), and whether the fv -it suffix immediately follows the verb root and immediately precedes any inflectional morphology, as in 5: (5) ag-it-aret root-fv-infl ‘he may have agitated …’ after the fv tokens were located, a randomizing function was called shuffling the fv data. results files were generated containing a random sample of 15 fv tokens for the top nine most frequent fv verb lexemes (including the token, context, and author and source for each). 2.2. procedure and analysis i consider three levels of linguistic abstraction in my analysis – morphology, syntax, and semantics – and how those levels interact with respect to latin fvs and their base verbs. the proceeding subsections outline the specific morphological, syntactic, and semantic features noted for each verb token. 3 this figure is a best approximation and will be finalized as the corpus is further developed. 6 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 morphological properties two morphological characteristics were noted in the annotation of each fv and base verb token: grammatical voice, and whether the token was a base or derived (i.e. fv) form. the grammatical voice of the token was determined by the presence of active or passive verbal morphology, which is largely systematic and predictable in latin4, in conjunction with recognized syntactic properties of passivization, such as demotion of the actor from the nominative case to an oblique (typically ablative, or a prepositional phrase if syntactically realized), and promotion of the undergoer from the accusative case to nominative (dixon and aikhenvald 2011). additionally, as noted in section 2.1, the derived form was determined based on the transparent morphological structure of the fv relative to its base verb form (i.e., the presence of the -it fv morpheme and first conjugation morphology). syntactic properties in addition to word-level properties, i consider three syntactic properties: verb valency, the syntactic frame (and syntactic restrictions), and the number of syntactically realized protoagents and proto-patients. verb valency was determined by the number of arguments present in the sentence or clause that are semantically necessary to the proposition expressed by that verb. for example, in example 6 the arguments needed to ‘complete’ the verbal action are hic ‘that man’ (agent), patriam ‘country’ (theme), and auro ‘for gold’ (asset), indicating that this verb token vendidit has a valency of three. (6) vendidit hic aur-o patri-am sell that.nom gold-dat country-acc ‘that traitor sold his country for gold’ (verg. a. 6.621) syntactic frames were modeled on those utilized in the verbnet (kipper-schuler 2005) verb lexicon and were considered in the present analysis with the hypothesis that syntactic frames can correlate with semantic class membership (levin 1993), in the same sense as verbnet. an example syntactic frame might be [np[nom] v np[acc] np[dat]] for 6 above (the syntactic frames’ word 4 with the exception of deponent and semideponent verbs, which are morphologically passive but syntactically and semantically active. 7 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 order reflects english’s for the purpose of analysis, mirroring the syntactic representations in verbnet, but with the awareness that word order is flexible in latin and is not used to track grammatical relations). in cases where the syntactic subject had been ‘dropped’ (facilitated by subject agreement morphology on the verb), a pro constituent was provided in the syntactic frame. additionally, restrictions were included on each syntactic constituent as subcategorization information, reflecting information such as: • the specific case argument nps appeared in (e.g. np[nom] for nominative, np[dat] dative). • voice for passivized verbs (e.g. np[nom] v.pass). • the specific preposition heading pps (e.g. pp[ex] for ‘out of…’, pp[ad] for ‘toward…’). • the sentential complement type (e.g. s[inf] for a bare infinitive, s[cx] for a syntactic construction). finally, quantitative figures were collected for the number of syntactically realized protoagents and proto-patients (introduced below) for each fv and base verb lexeme, since the overt realization of proto-agents and proto-patients can be seen to reflect the degree of transitivity of the verb – an overt (individuated) subject is typically more agentive (and nonindividuated as less agentive), and an overt object is typically more physically affected (and nonindividuated is less physically-affected) (hopper and thompson 1980:252). semantic properties at the highest level of abstraction, in the sense that the morphological and syntactic properties tend to be physically instantiated and directly observable, certain semantic properties were considered: semantic role array (and selectional restrictions), verbnet semantic class, and the number of agentive proto-agents and physically affected (or having undergone changes in state) proto-patients. i assume the same set of semantic roles as those used in verbnet (kipper-schuler 2005), and in order to facilitate any potential semantic generalizations, these fine-grained semantic roles (e.g. agent, stimulus, experiencer, theme, patient, etc.) were also considered in terms of their classification into dowty’s (1989) two ‘macro-roles’: proto-agent (pag) (expressing volition, sentience, causation, etc.) and proto-patient (ppt) (change of state, causally affected, etc.). furthermore, for each semantic role certain selectional restrictions were noted in a similar fashion 8 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 to verbnet, such as whether the role(s) require a +/animate argument. however, i consider one additional selectional restriction as an additional measure of overall transitivity: whether the verbal arguments are +/referential, in the strict semantic sense that the argument is + referential if it denotes an individual or property in the world (kearns 2011:2). this additional semantic feature can therefore be seen as reflecting the degree to which that argument is truly agentive (pag) or physically affected (ppt) (hopper and thompson 1980:252). in addition, i mapped each verb token to an existing verbnet semantic class (vnc), in order to generalize across several related and potential interpretations of that particular token, as well as to assist in the determination of the token’s semantic role array. in certain ambiguous cases, the possible english senses of the token were located in the proposition bank (propbank) (palmer et al. 2005) english framesets index, and the vnc was then determined through the mapping between propbank and verbnet, facilitated by the unified verb index5. finally, quantitative figures were calculated reflecting several semantic effects of the fv derivation, such as the number of agentive pags, physically affected ppts, +/referential pags, +/animate ppts, etc. 3. results results of the linguistic analysis demonstrate that while some linguistic features are unhelpful in attempting to characterize the latin fv verb, such as the number of valent members and passive or active morphology, other syntactic and semantic properties do appear to be helpful, albeit to varying degrees and depending on whether a derivational or usage-based perspective is assumed. in section 3.1 i discuss possible derivational generalizations, followed by a discussion of the usagebased generalizations in section 3.2. 3.1. derivational characterizations from a strictly derivational perspective, no clear semantic generalization was able to be formed in regard to the semantic class membership of latin fv verbs. however, the application of the fv -it suffix does appear to affect particular semantic properties in a general manner, primarily concerning the pag, ppt, and potentially the transitivity of the derived fv form. 5 the unified verb index merges information from four natural language processing projects: verbnet, propbank, framenet, and ontonotes sense groupings. online: https://verbs.colorado.edu/verb-index/index.php. 9 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 effects on pags the fv derivation seems to disfavor referential pags in comparison to their base verb semantics, however this generalization is slightly variable. for example, the derived fv verb form cit (represented by its root) was observed to occur only twice with a referential pag and 13 times with a nonreferential pag compared to its base verb ci(e), whose occurrence is more evenly distributed between referential and nonreferential pags: nine times with a referential pag, and six times with a nonreferential pag (a full table reflecting these figures is provided in table 1 in the appendix). examples 7 and 8 illustrate this: (7) latin-us dux [ … ] proeli-um ciet latin-nom leader.nom [ … ] battle-acc urge ‘the latin leader [not the least discouraged by his wounds] urged on the fighting’ (liv. 2.19) (8) labor optim-os citat labor.nom best-acc call ‘work calls for the best men’ (sen. prov. 1.5) in 7, livy is recounting in his history of rome an intense battle among several soldiers, and our example occurs just as the latin leader is injured and withdraws from the battle. the pag latinus dux ‘latin leader’ is syntactically realized, clearly referential, and functions as the subject to the base verb cie. in contrast, although also syntactically realized, the pag in 8 is largely nonreferential, in that it expresses a highly abstract concept ‘labor; work; toil’. the pags of the derived fv forms tend to be in general nonreferential and less potent in their agency, indicating a lower degree of transitivity (hopper and thompson 1980:252). therefore, a preliminary conclusion concerning the derivational effects of the fv suffix might be the reduction in agency and transitivity of the derived lexeme, considering approximately 73% of all fv forms sampled occur with nonreferential pags (cf. table 1) in contrast with an approximately even distribution of base verb forms occurring with referential and nonreferential pags. effects on ppts fv derivational effects on ppts were similar to those on pags; specifically, the derived verbs seem to prefer nonreferential ppts, with approximately 40% of fvs occurring with nonreferential ppts as opposed to approximately 34% of base forms occurring with nonreferential ppts (cf. 10 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 table 1). however, derivational effects on ppts appear to be more uniform, and in general seem not to result in a change of state or take a physically affected ppt. for example, the fv habit never occurred with referential ppts or physically affected ppts, compared to its base verb hab, which was observed to occur eight times with a physically affected ppt and 11 times with a referential ppt. similarly, the fv vendit was observed to occur only three times with a physically affected ppt and once with a referential ppt, whereas its base verb vend occurred 14 times with a physically affected ppt, and 13 times with a referential ppt. these effects of the fv derivation on ppts are also apparent between the fv and base verb pair cogit and cog, as shown in 9 and 10: (9) eos ipsos [ … ] ced-ere in tutm coegit pro.masc.pl.acc those-masc.pl.acc [ … ] withdraw-inf into safety force ‘he forced/compelled them to withdraw to a place of safety’ (liv. 2.10) (10) nihil de resistendo cogitabat nothing.acc concerning resisting consider ‘they (inferred attinian army) never thought of making resistance’ (caes. civ. 2.34) it is important to note that, similar to the fv habit, the fv cogit never takes a physically affected or referential ppt, whereas its base cog almost exclusively does. example 9 occurs as livy recounts an attack on rome, and how one individual, horatius cocles, had been able to prevent the onslaught. the ppt ipsos eos ‘those individuals (masculine plural)’ is referential, referring to a group of individuals (roman citizens), and evidently undergoes a change of state, being the object in an object control construction in which the ppt is the syntactic object of the control verb (coegit ‘force’) and the syntactic subject of the subordinate infinitival cedere ‘withdraw’. furthermore, the subordinate action of ‘withdrawing’ involves a change in location, further highlighting the change of state of the ppt ipsos eos. the fv cogitabat ‘consider’ in 10 takes a nonreferential ppt (nihil de resistendo ‘nothing concerning resisting’), and is neither physically affect nor undergoes a change of state. analogous to the prior observation concerning the fv derivation’s effects on pags, the tendency of fvs to take nonreferential, nonphysically affected ppts suggests reduction in transitivity to be a result of the application of the fv suffix, as well as a reduction in agency of the pag. 11 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 3.2. usage-based characterizations in contrast to a derivational (or source-oriented (bybee 2001:126)) approach, which attempts to specify the base verb form and a single operation with a single set of conditions (croft and cruse 2004:301) (in this case, the application of the fv -it suffix and the attempt to predict and generalize over its semantic effects), fvs can be investigated in terms of “conditions” on the fv form only (becker and fainleib (2009:3). when viewed in terms of a usage-based, or product-oriented (bybee 2001:126), the semantic classification of fv tokens appear to be influenced by their syntactic frames, semantic role array, and particular lexemes appearing in their argument complementation. frames, roles, and vnc membership examples 11 and 12 illustrate the relationship between vnc membership and particular syntactic frames and semantic roles: (11) cum se cogitat esse pi-um adv himself.acc consider be.inf pious-acc ‘when he knows that he is pious’ (lit. ‘considers himself to be..’) (catul. 76) (12) sed [ … ] cogitavit [ … ] reg-es barbar-os incit-are conj [ … ] plan [ … ] ruler-acc foreign-acc incite-inf ‘but he planned [from the start] […] to incite foreign rulers’ (cic. att. 8.11) the frame [pro v np[acc] np[inf]] in 11 exclusively occurred with the ‘consider, know’ interpretation of the fv cogit correlating with the consider-29.9 class, as matched to verbnet. furthermore, the semantic role theme also only occurred with the consider-29.9 class, indicating a relation between the semantic role(s) of the fv verb’s core arguments and its semantic class membership. in comparison, the syntactic frame [pro v s[inf]] exhibited in 12 exclusively occurred with the ‘plan, intend’ interpretation of cogit, correlating with the wish62 semantic class (its closest match in verbnet). lexical complementation and vnc membership another structural feature that tended to correlate with fv’s vnc membership concerns particular lexemes appearing in the complementation of the verb, as demonstrated in 13 below: 12 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 (13) et defendend-ae urb-is consili-a agitaba-ntur conj defending-gen city-gen plans-nom discuss-pass ‘plans of defending the city were discussed’ (liv. 10.21) semantically, the fv agit in example 13 might belong to nine6 possible vncs. however, the lexeme consilia ‘plans, counsel’ (and a related lexeme res ‘thing, event, fact’) only occurred with the discuss interpretation (and vnc classification) of agit. a similar relationship was observed among the fv vendit sample: when the ppt being ‘sold’ is the reflexive pronoun se ‘oneself’, coindexed with the pag, the interpretation of agit is exclusively ingratiate – a metaphorical extension, considering the literal translation as ‘sell oneself’. it is interesting to note that these highly metaphorical, idiomatic interpretations, and particular lexical complementation patterns, tend to occur with lower frequency fv lexemes (e.g. flagit (90 token frequency), vendit (54), and dict (46)), compared to higher frequency fvs such as cogit (748); this will be explored in future work. 4. concluding remarks this study considered certain morphological, syntactic, and semantic properties of latin fv verbs from two synchronic perspectives: derivational (source-oriented) and usage-based (productoriented). in attempting to characterize the fv as a derivational process, the fv -it suffix appears to affect whether the derived verb’s pag and ppt are referential, tending to take non-referential core arguments in relation to their base verb forms, and unaffected ppts. this can be viewed as a reduction transitivity and agency (hopper and thompson 1980), which follows from fvs’ denominal origins (greenough et al. 2001:152) with denominal verbs typically denoting a state (viti 2012:7) and low transitivity. the fv derivation can therefore be seen as a potential stativizing 6 the possible vncs to which the fv agit might belong, as identified in this study, include: 1. discuss (no verbnet equivalent, sense from propbank) 2. amuse-31.1 3. establish-55.5-1 4. conduct (no verbnet equivalent, sense from propbank) 5. handle (ibid.) 6. force-59 7. judgement-33 8. push-12-1-1 9. risk-94 13 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 process as reflected by its argument selection and overall reduction in transitivity. in contrast, a usage-based, product-oriented analysis reveals that specific syntactic frames, semantic roles, and lexical complementation allows fv lexemes to be clustered in terms of prototypical semantic classes, providing a more straightforward characterization of latin fv verbs and highlights the interconnectedness of morphology, syntax, and semantics. references becker, michael, and lena fainleib. 2009. the naturalness of product-oriented generalizations. amherst, university of massachusetts amherst, ms. online: http://www.phonologist.org/hebrewplurals/. bybee, joan. 2001. phonology and language use. cambridge: cambridge university press. crane, gregory r. (ed.) 2005. perseus digital library. medford, tufts university. online: http://www.perseus.tufts.edu. croft, william, and d. alan cruse. 2004. cognitive linguistics. cambridge: cambridge university press. dixon, r. m. w., and alexandra aikhenvald. 2011. a typology of argument-determined constructions. in language at large: essays on syntax and semantics. leiden: brill. dowty, david r. 1989. on the semantic content of the notion ‘thematic role’. in gennaro chierchia; barbara h. partee; and raymond turner. (eds.) properties, types and meaning, ii. dordrecht: kluwer academic publishers. greenough, j. b.; g. l. kittredge; a. a. howard; and benj. l. d'ooge. (eds.) 2001. allen and greenough’s new latin grammar. newburyport: focus publishing, r. pullins & company. hopper, paul j., and sandra a. thompson. 1980. transitivity in grammar and discourse. language 56.251-299. kearns, kate. 2011. semantics. new york: palgrave macmillan. kipper schuler, karin. 2005. verbnet: a broad-coverage, comprehensive verb lexicon. philadelphia: university of pennsylvania dissertation. levin, beth. 1993. english verb classes and alternations: a preliminary investigation. chicago: university of chicago press. palmer, martha; daniel gildea; and paul kingsbury. 2005. the proposition bank: an annotated corpus of semantic roles. computational linguistics 31. 14 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 solodow, j.b. 2010. latin alive: the survival of latin in english and the romance languages. cambridge: cambridge university press. viti, carlotta. 2012. the use of frequentative verbs in early latin. proceedings of the xvi colloquium internationale linguisticae latinae, uppsala. 15 desjardins: a corpus-based linguistic analysis of latin frequentative verbs published by cu scholar, 2019 appendix – table 1. fv and base verb quantitative figures fv base token freq. vncs overt pags overt ppts transitive/ agentive pags phys. affected ppts pags +/ref. ppts +/ref. agit 362 9 2 14 0 2 1 / 14 2 / 13 ag 6 3 12 2 4 7 / 3 5 / 10 cit 427 3 4 14 10 2 2 / 13 12 / 3 ci(e) 2 9 15 11 15 9 / 6 11 / 4 cogit 748 2 6 13 10 0 4 / 11 0 cog 1 2 9 11 15 4 / 11 10 / 5 concit 278 1 4 14 8 15 5 / 10 11 / 4 conci(e) 1 9 15 15 15 6 / 9 10 / 5 dict 46 1 6 12 15 15 3 / 12 8 / 7 dic 4 6 15 11 4 7 / 8 9 / 6 excit 311 3 7 14 12 15 5 / 10 8 / 7 exci(e) 4 9 15 13 15 4 / 11 9 / 6 habit 269 1 9 0 0 0 9 / 6 0 hab(e) 2 5 14 5 8 5 / 10 11 / 4 flagit 90 2 4 11 7 1 4 / 11 9 / 6 flagr 3 11 0 0 0 10 / 5 0 vendit 54 2 3 14 3 3 3 / 12 1 / 14 vend 1 12 14 12 14 12 / 3 13 / 2 appendix – roman works and authors referenced in this paper caesar, gaius julius. c. 100 bce – 44 bce. de bello civili. catullus, gaius valerius. c. 84 bce – 54 bce. carmina. cicero, marcus tullius. c. 106 bce – 43 bce. epistulae ad atticum. livius (livy), titus. c. 59 bce – 17 ce. ab urbe condita. seneca, lucius annaeus. c. 4 bce – 65 ce. ad lucilium epistulae morales and de providentia. vergilius (vergil) maro, publius. c. 70 bce – 19 bce. aeneidos. 16 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/1 doi: http://dx.doi.org/10.33011/cril.24.1.1 colorado research in linguistics 6-2019 a corpus-based linguistic analysis of latin frequentative verbs jared desjardins recommended citation a corpus-based linguistic analysis of latin frequentative verbs cover page footnote microsoft word desjardins-cril2019-final.docx female-to-male transsexuals and gay-sounding voices: a pilot study colorado research in linguistics. june 2010. vol. 22. boulder: university of colorado. © 2010 by lal zimman. female-to-male transsexuals and gay-sounding voices: a pilot study* lal zimman university of colorado a great deal of work has now been published on the perception of men’s sexual orientation on the basis of phonetic characteristics. in this paper, i present a pilot study focusing on a population that sheds new light on this topic: female-to-male transsexuals. as individuals who were raised as girls but self-identify as men, trans men (as they are also called) are often perceived as gay-sounding after undergoing the drop in vocal pitch that is typically brought on by testosterone therapy. using recordings of read speech from three trans men and five non-trans men who were each rated as gayor straight-sounding by listener subjects, the analysis presented here shows that trans men are perceived in much the same way as gay-sounding non-trans men, despite a number of differences in the acoustic features of their voices. ultimately these findings lend credence to the notion that there is no single gay-sounding phonetic style, but rather multiple styles that are lumped together perceptually as gay-sounding on the basis of their deviation from norms for straight-sounding voices. 1. introduction as the study of language and sexuality has become an established subdiscipline within sociolinguistics over the past two decades, a number of linguists have taken an interest in the question of whether sexual orientation can be detected on the basis of particular phonetic features or styles, particularly among male speakers (gaudio 1994; levon 2007; linville 1998; munson, jefferson and mcdonald 2006; munson et al. 2006; munson 2007; pierrehumbert, et al. 2004; podesva, roberts and campbell-kibler 2001; podesva 2007; smyth and rogers 2002; and smyth, jacobs and rogers 2003).1 while one of the most basic questions explored in this literature has been whether listeners can accurately judge male speakers as gay or straight based on voice alone, these authors have also sought to uncover the precise phonetic features that correlate with the perception of a man’s voice as “gay-sounding.” this paper is a revised version of the author’s preliminary examination for the phd in linguistics at the university of colorado, boulder. thanks are owed to rebecca scarborough for her help in designing this research, and to steven duman, joshua raclaw, richard sandoval, susanne stadlbauer, and an anonymous cril reviewer for feedback during the revision of this work. 1 a few studies (waksler 2001; pierrehumbert et al. 2004; munson et al. 2006; munson, jefferson & mcdonald 2006) have also examined the voices of lesbian women and/or women perceived as lesbian-sounding. the findings of these studies are important, but beyond the scope of the present paper. 1 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 2 in this paper, i focus on these same perceptual and acoustic questions, but i also provide new insight through the introduction of a group of speakers that are almost completely absent from the sociocultural linguistic literature: female-tomale transsexuals. female-to-male transsexuals, who are also called trans men, are individuals who are assigned to a female gender role and raised as girls, but who come to identify as men at some point later in life. trans men make for an especially interesting community for sociophonetic study because of their unique set of experiences with the biological and socialized aspects of the voice. while trans men typically have female-sounding voices before their shift from a female social role to a male one, these individuals generally come to be heard as malesounding over the course of their gender role transition. this is in large part due to the fact that many trans men make use of testosterone for hormone replacement therapy, which results in a marked drop in vocal pitch just as it does during typical male puberty. in fact, as i show in later sections of this paper, trans men’s voices are indistinguishable from the voices of other men when it comes to fundamental frequency. on the other hand, many of the differences between men’s and women’s voices have been shown to be learned during childhood rather than determined by the biological differentiation that arises during adolescence. and since trans men are expected to grow up as girls and later become women, their experiences with socialization are different from those had by men who are raised as boys. the questions driving the research described in this paper have to do with the consequences of this mixture of biological and social factors that is characteristic among trans men. because some authors studying gay-sounding voices have suggested that boys who acquire “feminine” phonetic traits during childhood might come to sound gay as adults (smyth and rogers 2002; renn 2002), my goal in undertaking this work has been to explore whether trans men would be described as gay-sounding by listeners in a perceptual experiment like those conducted by other authors. if trans men’s voices do indeed tend to be gaysounding, i also aim to discover which vocal features might explain this perception, and whether the acoustic characteristics of trans men’s voices are the same as those found among gay-sounding non-trans men. i begin this paper with an overview of the findings of previous studies on gay-sounding men’s voices, which provide a starting point for my own analysis. notably, the findings of these studies have often been contradictory, making it difficult to construct a unified model of gay-sounding men’s voices. rather than presenting a challenge to be overcome, however, i argue that the differences in these findings likely reflect real-life diversity among gay-sounding speakers. in other words, different studies have reached different conclusions because there are in fact multiple phonetic styles that might be interpreted as gay-sounding. in section 3, i discuss my own research comparing the voices of trans men to the voices of both gay-sounding and straight-sounding non-trans men. what my findings show is that trans men are indeed perceived as gay-sounding: members of this group were rated by listeners in the same way as the gay-sounding non2 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 3 trans men. however, my phonetic analysis reveals a number of differences between these groups when it comes to acoustic measurements. in the discussion section i ultimately argue that these findings support the idea, introduced by zwicky (1997), that different phonetic styles are lumped together as gay-sounding by listeners simply because they deviate from a straight-sounding norm. i conclude with reflections on how the linguistic practices of trans men can enhance our understanding of gay-sounding voices more generally, and with directions for future research. 2. previous research a number of studies, such as those carried out by gaudio (1994), linville (1998), and smyth, jacobs and rogers (2003), have compared speakers’ selfidentified sexual orientation to listeners’ perception of these same individuals as straight or gay on the basis of read speech. consistently, such work has shown that listeners are able to identify speakers’ sexual orientations at better than chance rates. at the same time, however, each of these authors have noted that the correlations they have uncovered are imperfect, simply because not all gay men sound gay and a few straight men do. nevertheless, this work suggests there is a salient socio-perceptual category for “gay-sounding” voices. having shown that such a grouping exists on the perceptual level, these authors and others have focused on uncovering the acoustic characteristics that correlate with the categorization of a particular voice as gayor straight-sounding. as i mentioned in the previous section, it is difficult to synthesize the findings of research on gay-sounding voices because of the way different studies have sometimes reached contradictory conclusions. for example, most research that has investigated speakers’ mean fundamental frequency have shown no difference between gayand straight-sounding men on this measure (gaudio 1994; linville 1998; podesva, roberts and campbell-kibler 2001; smyth and rogers 2002). however, munson et al.’s (2006) study, which analyzed words produced in isolation rather than connected speech, found that the gay-sounding men in their study did have higher fundamental frequency than the straightsounding men. the same study also found that gay-sounding men had higher mean f1 and f2 than straight-sounding men, with munson (2007) further confirming the significance of mean f1. however, all of the other studies that have compared mean f1 and f2 across gayand straight-sounding speakers show no significant differences (linville 1998, smyth and rogers, 2002; pierrehumbert et al. 2004). vowel duration also seems to play some role in the perception of men’s sexual orientation, but it isn’t clear whether this difference is found only in certain vowels (podesva, roberts and campbell-kibler 2001; smyth and rogers 2002), in all vowels (munson et al. 2006), or potentially not at all (pierrehumbert et al. 2004). crucially, whether a variable correlates with listeners’ perception of a voice as gay-sounding seems to depend in part on the other variables present in 3 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 4 the speakers’ style: levon (2007) found that decreasing a gay speaker’s pitch range through digital manipulation lead to reduced gayness ratings by listeners, but increasing a straight speaker’s pitch range had no significant effect on how his voice was rated. some of the variation that shows up in these findings is no doubt due to differences in data collection methods, which have included individual words (munson jefferson and mcdonald 2006; munson et al. 2006; munson 2007) and connected read speech (gaudio 1994; linville 1998; pierrehumbert et al. 2004; smyth and rogers 2002) and spontaneous speech (smyth, jacobs and rogers 2003) produced in laboratory conditions, as well as an unscripted radio broadcast (podesva, roberts and campbell-kibler 2001). however, there is another very important potential explanation, which is that speakers perceived as gay-sounding are probably not all using the same phonetic style. as i mentioned in the previous section, this is the argument advanced by zwicky (1997). specifically, he says that there is unlikely to be a single set of characteristics that can be delineated as the gay-sounding style (also see podesva, roberts and campbell-kibler 2001 for a highly nuanced treatment of this idea). instead, as zwicky suggests, virtually any deviation from ways of talking associated with heterosexual masculinity – whether by virtue of a higher pitch, higher first and second formants, greater duration for vowels and/or consonants, or any of the other of numerous features that have been investigated – can be interpreted as indexing gay identity. this idea is highly intuitive, given the pervasive cultural discourse that equates any kind of gender non-normativity, particularly among men, with homosexuality (discussed in detail by gaudio 1994). there are also empirical findings to support zwicky’s argument. gordon (2008) presents an analysis that compares the same speakers delivering gay-sounding and straight-sounding readings of the same passage. gordon found that speakers made use of a wide variety of styles in their gay-sounding guises, each characterized by different marked phonetic variants, while their straight-sounding guises were much more similar to one another’s. what these speakers’ gay-sounding readings had in common, then, was their deviation from a more homogenous straight-sounding style. although authors have often reached different conclusions regarding which acoustic features are salient in the perception of sexual orientation, many researchers working on this topic have pointed out the similarities between the speech of gay-sounding men and that of women. particularly in the work of smyth and his colleagues (especially smyth and rogers 2002), gay-sounding voices have been characterized as a mixture of features typically associated with men’s voices, such as a relatively low mean f0, with other features typically associated with women’s voices, such as relatively longer sibilants or vowels that are articulated closer to the periphery of the vowel space. given these similarities, some authors have suggested that men who reject or fail to conform to heteronormative masculinity are more apt to sound gay than men with more conventional and ideologically unmarked enactments of gender (renn 2002; smyth and rogers 2002). more specifically, gender socialization during 4 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 5 childhood is presented in this work as a likely source for gay-sounding voices among men, given that children are known to acquire sociophonetic markers of gender quite early in life (e.g. sachs 1975). for instance, smyth and rogers (2002) argue that despite getting similar linguistic input from adults, some boys may engage in selective intake by orienting more strongly to women speakers as linguistic role models rather than men.2 considering what we know about the acquisition of other forms of sociolinguistic distinction (such as regional dialect) and the way that adolescents negotiate their role in the heterosexual marketplace as part of a peer-based social order (eckert 2003), young speakers’ age cohorts probably also play a considerable role in the socialization of gendered phonetic traits. based on the arguments presented by these authors, it is worth considering the possibility that men who grow up orienting to the norms for women speakers in their communities, rather than men, will tend to be judged as gay-sounding in adulthood. the present study addresses this question from a rather unusual angle: by focusing on the voices of trans men. as i mentioned in the introduction, trans men very often make use of testosterone therapy as part of their transition from a female gender role to a male one, which generally results in a great deal of physiological masculinization, including changes in the larynx. one study of trans men’s voices, which appears to be the only of its kind (described in both van borsel et al. 2000 and in adler and van borsel 2006), found that the two individuals studied experienced a significant decrease in mean f0 and in f0 range, which put them within a normative male range during the first year of testosterone therapy. on the other hand, testosterone has no apparent effect on the many phonetic cues for speaker gender that are learned during language socialization, such as differences in segment duration or vowel quality (see simpson 2009 for a review). of course, given that trans men are raised in a female gender role, their experiences with childhood language socialization are markedly different from most men’s. if trans men do differ from most other men in terms of socially-learned gendered phonetic traits, and if these speakers are perceived as gay-sounding men, then the unique experiences of members of this group would seem to provide evidence that childhood gender socialization can be a significant factor in predicting whether a man will be perceived as gayor straight-sounding, at least for some speakers. in order to explore this issue, the remainder of this paper is devoted to a comparison of men from three groups: trans men (hereafter tm), non-trans men with gay-sounding voices (gsm), and non-trans men with straight sounding voices (ssm). in this space i focus on two questions: first, how are the voices of tm perceived, compared to gsm and ssm? second, what are some of the acoustic similarities and differences between members of these groups? before 2 of course, it might just as easily be the straight-sounding boys who are engaging in selective intake by orienting only to men rather than also paying attention to women, or that all children are engaging in some kind of selective intake in choosing their speaker role-models. 5 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 6 answering these questions, section 3 will describe the methods for data collection and analysis used in this project. 3. methods 3.1. data collection in order to test the questions presented above, speakers were recruited through previous research contacts and the researcher’s own extended social network at the university of colorado and in san francisco, ca. none of the participants were aware of the topic of the investigation before being recorded beyond the fact that they were participating in a study about “how men talk.” three speakers were recorded from each of the three groups under investigation: tm (trans men), gsm (gay-sounding non-trans men), and ssm (straightsounding non-trans men). the gsm and ssm speakers were initially selected on the basis of whether i perceived them to be gay-sounding or straight-sounding, but these perceptions were then checked against listener ratings (see section 3.2 below). speakers were from urban or suburban areas in the western us and were between the age of 20 and 27, with the exception of one 47 year old speaker in the gsm group. while the non-trans men each identified as either gay or straight, all of the trans men identified with broader and potentially more fluid sexuality labels, such as bisexual, pansexual, and/or queer. following the methodology described by smyth, jacobs and rogers (2003), speakers were recorded while reading two passages: the rainbow passage (fairbanks 1960), which is an historical and scientific overview of rainbows, and the fire passage (crist 1997), which is a dramatic narrative about a building fire. however, the final analysis presented in this paper includes only the fire passage (the text of this passage can be found in appendix a). this choice was motivated primarily by smyth, jacobs and rogers’ finding that the scientific rainbow passage tended to evoke inflated gayness ratings for speakers who were perceived as straight in other contexts. additionally, these authors found no significant differences between the dramatic read passage and a spontaneous spoken passage, suggesting that the dramatic passage is more representative of speakers’ more naturalistic speaking styles.3 in addition, technical problems with the computer used for recording meant that a few speakers had to reread the scientific passage, which had a clearly audible effect on the speed at which they read; obviously, this would problematize the comparison of segment duration. while read speech is known to differ from naturally-occurring discourse in a number of ways and thus limits the generalizability of this study, read speech was chosen to facilitate the perceptual experiment described in section 3.2 as well as providing easily 3 two volunteers for this study also pointed out the symbolic significance of rainbows in the gay community. although smyth, jacobs and rogers assume that genre is the only factor at work here, it could be that the topic also influences listeners’ judgments or even speakers’ production. 6 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 7 comparable data for acoustic analysis. however, these findings should be seen as a starting point that will guide future studies that make use of interactional language data (see section 5). recordings were made with a samson c03u usb multi-pattern condenser microphone and digitized at 48,000 hz using the audacity recording software program (audacity development team 2008). 3.2. listener evaluation identical segments of approximately 30 seconds were extracted from each speaker’s reading of the fire passage. eight listener subjects, all native speakers of american english, were then recruited to evaluate these clips via an online survey. the survey presented each audio clip along with sets of binary adjectives from which listeners were instructed to choose; for example, they were asked to rate how tall or short each speaker sounded on a scale of 1 to 5.4 listeners were instructed to play the audio clips and mark each speaker’s characteristics according to their best guess, but there was also an option to choose “no clue” to signify that the listener had no guess whatsoever as to a particular characteristic. listeners were not instructed on the purpose of the experiment, nor that collecting the gay versus straight ratings were the primary purpose of the study. discussion of the other social characteristics listeners rated is beyond the scope of the present analysis. on the basis of listener perceptions, one speaker from the gsm group was excluded – despite my perception of him as gay-sounding, listeners consistently perceived him as a straight-sounding speaker (i.e. his gayness ratings were not significantly different from the straight men in this study). because the goal was to compare the voices of straight-sounding straight men and gay-sounding gay men (rather than defining groups primarily on self-identification), only 2 speakers from the gsm group were included for analysis. 3.3. acoustic analysis each thirty second clip that was played for listener subjects was also subjected to acoustic analysis using the praat software package (boersma and weenink 2008). the features chosen for this analysis were selected on the basis of previous studies’ findings and included the following measures: 1. voiceless sibilant consonants (20 tokens of /s/, 1 token of /ʃ/) a. mean duration b. mean frequency at peak amplitude c. mean center of gravity 4 other traits, in addition to gay versus straight and masculine versus feminine, included young versus old, rude versus polite, and short versus tall. 7 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 8 2. vowels a. mean f0 across 11 stressed vowels b. f0 range across 11 stressed vowels c. mean f1 and f2 across 11 stressed vowels d. f1 and f2 of /æ/ (3 stressed tokens) and /ɛ/ (2 stressed tokens) given the stereotype of the “lisping” gay man, it is unsurprising that researchers of gay-sounding voices have often directed their attention to sibilant consonants. indeed, variation in the properties of /s/ has been consistently shown to correlate with the perception of men’s voices as gayor straight-sounding. many researchers have investigated the duration of these sounds (linville 1998; podesva, roberts and campbell-kibler 2001; smyth and rogers 2002; though see also levon 2006, 2007), but some have also focused on the acoustic qualities of these segments themselves (linville 1998; munson et al. 2006; munson 2007). sibilant consonants, like other fricatives, are characterized by high-frequency aperiodic energy. fricatives can be distinguished from one another by which frequencies are most prominent in terms of amplitude. that is, while /s/ tends to have relatively high-amplitude energy at around 8,000 hz, the highest-amplitude energy in /ʃ/ tends to be closer to 4,000 hz (johnson 1997:130). thus, one measure that has been used in investigations of sibilants in general, and in gaysounding sibilants in particular, has been the frequency of the sound at peak amplitude (linville 1998). similar information can be gathered through the measurement of the center of gravity of /s/, which provides a holistic view of which frequencies have the highest amplitude within a sound. another measure that has received attention is spectral skew (munson et al. 2006; munson 2007), which refers to whether the majority of acoustic energy is located in the higher frequencies of the sound or the lower frequencies, but this particular measure was not used in the present study. in order to compare the sibilant consonants of the speakers in this study, i identified the instances of /s/ (n = 20) and /ʃ/ (n = 1) in the thirty second clips played for listeners. i then measured each token’s length, generated a spectral slice for the token,5 from which the peak frequency was identified visually, and finally generated center of gravity measurements using praat’s automated moments analysis function, which were checked against visual examinations of the spectra. because the 20 tokens of /s/ appeared in identical phonemic contexts, comparisons across speakers used the mean values of these measurements (i.e. mean duration, mean center of gravity, etc.). vowel quality has also been consistently shown to influence the perception of sexual orientation. although men with gay-sounding voices have not usually been found to have overall higher mean formants than straight sounding men, some research has turned up differences in the quality of individual vowels in terms of either f1 or f2 (smyth and rogers 2002; 5 a spectral slice provides a visual representation of the relationship between frequency and amplitude within a sound at a given point in time. 8 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 9 pierrehumbert et al. 2004; munson et al. 2006). in this study, i discuss two of these vowels: /æ/ and /ɛ/. i analyzed three stressed instances of /æ/ and two stressed instances of /ɛ/ from each 30-second sample. in addition to the five tokens of /æ/ and /ɛ/, six other stressed vowels were included in the analysis in order to have a broader range of data from which to calculate f0 measurements, including three tokens of /i/, two tokens of /a/, and a single token of /ʌ/. each of these vowels was measured for duration; maximum, minimum, and mean f0 across the entire token; mean f1 and f2 across the entire token; and f1 and f2 at two points in the vowel, approximately 1/3 and 2/3 of the way through the segment (estimated visually). 4. results 4.1. perceptual results before discussing the acoustic findings of this project, it is important to establish which speakers were perceived as gay-sounding and which were perceived as straight-sounding. aside from the gsm speaker who was eliminated from the sample (see section 3.2 above), listener evaluations correlated strongly with my preliminary groupings of men as gayor straight-sounding (see appendix b for listener ratings). a one-way anova test6 showed that speakers’ average numerical rating for gayness interacted significantly with the group in which they had been placed (ssm, gsm or tm). the group effects, as calculated by a post hoc (tukey hsd) test, can be seen in table 1. table 1: speaker grouping vs. gayness rating group comparison p-values ssm vs. gsm 0.0171 * ssm vs. tm 0.0313 * gsm vs. tm 0.5994 .. * = significant at .05 as the table shows, there was a highly significant difference between the gayness ratings given to the speakers in the ssm and gsm groups, confirming 6 statistical analyses were performed using the r project for statistical computing software (r development core team 2008). anova is a useful approach to the data in question because it allows for the comparison of data on a linear continuum instead of requiring the use of categorical variables like traditional statistical software for sociolinguistic analysis (i.e. varbrul). 9 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 10 that listeners shared my perception of the gsm speakers as gay-sounding and the ssm speakers as straight-sounding. notably, there was also a significant difference between the gayness ratings for speakers in the ssm and tm groups. however, there was no significant difference between the gsm and tm speakers. we are now in a position to answer one of the questions driving this research: how are the voices of trans men perceived, compared to gay-sounding and straight-sounding men? these perceptual data indicate that tm and gsm speakers are both perceived as significantly more gay-sounding than the ssm speakers. furthermore, the fact that there is no significant difference between the gayness ratings given to the gsm and tm groups suggests that these groups are lumped together perceptually as gay-sounding men, in contrast to the straightsounding speakers in this study.7 however, the second question to be addressed in this paper remains: are the voices of gsm and tm as similar acoustically as they are perceptually? 4.2. sibilants as i discussed in section 3, one set of measurements i took was of the voiceless sibilant consonants /s/ and /ʃ/, including duration, frequency at peak amplitude, and center of gravity. beginning with /s/, i calculated each speaker’s mean for duration, peak frequency, and center of gravity across 20 tokens. i then ran statistical tests that compared each of these means against three factors: first, speaker group (e.g. do speakers in the gsm and/or tm category have a longer mean duration for /s/ length than those in the ssm group?); second, gayness rating (do speakers who were rated by listeners as more gay-sounding have a higher peak frequency than those who were rated as less gay-sounding?); and finally, masculinity rating (do speakers who were rated by listeners as less masculine have a higher center of gravity than those who were rated as more masculine?). based on these tests, the only statistically significant interaction was between group and center of gravity (p < 0.0012). specifically, the tm group had a significantly higher center of gravity in the distribution of energy in /s/ than either the ssm or gsm groups. the specific group interactions, as shown by post hoc analysis, can be seen in table 2. these results show that while there was no statistically significant difference between ssm and gsm speakers, there were significant differences between both the ssm and tm groups and between the gsm and tm groups. in this case, then, the voices of tm and gsm are not alike acoustically, despite their similarity perceptually. while the interaction of center of gravity and group was the only statistically significant result from this set of measurements, a few other results 7 listeners were explicitly told that all speakers were male in order to avoid the possibility that some speakers might be perceived as female. 10 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 11 were statistically suggestive (i.e. with p-values less than .1), namely, both the interaction between masculinity rating and sibilant duration (p < 0.0947) and the interaction between masculinity rating and peak frequency (p < 0.0766) approached significance. both of these factors thus deserve ongoing attention in future extensions of this research. table 2: speaker grouping vs. center of gravity for /s/ group comparison p-values ssm vs. gsm 0.1152 . ssm vs. tm 0.0043 * gsm vs. tm 0.0012 * * = significant at .05 although there was only one token of /ʃ/ for comparison, center of gravity again provided a statistically significant result. however, in this case it is the gsm group that stands apart from the other two rather than the tm group. table 3: speaker grouping vs. center of gravity for /ʃ/ group comparison p-values ssm vs. gsm 0.1040 . ssm vs. tm 0.3384 . gsm vs. tm 0.0234 * * = significant at .05 as table 3 shows, there is a significant difference between gsm and tm groups, and a difference that is nearly statistically suggestive between ssm and gsm groups, but no significant or suggestive difference between tm and ssm groups. again, the tm and gsm groups are acoustically different, despite being perceptually similar. interestingly, however, in this case it is the tm speakers who are like ssm speakers, whereas for /s/ it was the gsm group that resembled the ssm group. also of interest is the fact that the gsm’s center of gravity was significantly lower than the other two groups, when we might expect it to be higher. 11 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 12 4.3. fundamental frequency pitch was examined through measurements of fundamental frequency, but no significant or suggestive results were obtained. neither mean f0 nor f0 range correlated with speaker group, gayness rating, or masculinity rating. tests were also run on the f0 values for each of the individual stressed vowels examined in section 3.4, and again no significant results were obtained. in fact, the variation within groups was obviously greater than the variation across groups – the best example of this is the fact that the tm group contained both the speaker with the highest mean f0 and the speaker with the lowest mean f0 of all participants in the study. 4.4. vowel formants like a number of other studies of gay-sounding voices, the data i analyzed showed no significant differences in speakers’ overall mean first and second formants. as smyth and rogers (2002) have pointed out, this is one of the ways in which gay-sounding men’s voices are different from women’s. indeed, i found no differences between the mean vowel formants for gayand straight-sounding speakers, and this held true for both the gsm group and the tm group. however, as i mentioned above, a few studies have shown individual vowels to be especially likely to differ between gayand straight-sounding speakers. the current analysis produced similar findings. as i mentioned in section 2.3, i focused on the vowel quality of 3 stressed tokens of /æ/ and 2 stressed tokens of /ɛ/ that appeared in the spoken excerpt played for listener subjects. instead of taking the mean formant values for these vowels, i compared each token of /æ/ and /ɛ/ individually. the first instance of stressed /æ/ appears in the sentence “they must have been trapped,” (see appendix a). for this token, a one-way anova test showed that speakers’ mean f2 for this vowel interacted significantly with speaker group (p < 0.0152). specifically, the gsm group had a lower mean f2 than either the tm or ssm groups. the effects of each individual group, as shown by a post hoc test, are in table 4. in this case there is again a significant difference between the gsm and tm speakers (p < 0.0286), further demonstrating that these groups are acoustically different even as they are perceptually similar. however, this is the only one of the three tokens of /æ/ that showed statistically significant interaction between f2 and speaker group (none showed significant interaction with gayness rating or masculinity rating), making this finding tentative until further analysis is carried out. one other token of /æ/ did show a statistically suggestive interaction between mean f1 across this vowel and both speaker gayness rating (p < 0.0972) 12 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 13 and speaker masculinity rating (p < 0.0986). however, neither of the other two instances of /æ/ showed any significant or suggestive variation in f1 values. table 4: speaker grouping vs. f2 for /æ/ group comparison p-values ssm vs. gsm 0.0153 * ssm vs. tm 0.7619 . gsm vs. tm 0.0285 * * = significant at .05 the other vowel investigated in this analysis was /ɛ/. both instances of this vowel show a relationship between f1 and speaker gayness ratings as well as between f1 and speaker masculinity ratings. in the first token of /ɛ/, which appears in the sentence, “but as soon as i poked my head out, i smelled smoke,” there was a significant correlation between mean f1 for this vowel and gayness rating (p < 0.0432) as well as between f1 and masculinity rating (p < 0.0446). that is, speakers who were more gay-sounding (or less masculine-sounding) had relatively higher f1 values for /ɛ/. the second token of this vowel, from the sentence, “the ambulance guys had to put a splint on his leg,” showed the same pattern, but with only a statistically suggestive correlation between f1 and gayness rating (p < 0.0972) and between f1 and masculinity rating (p < 0.0986). the second token of /ɛ/ also showed a statistically suggestive correlation between f2 and gayness rating (p < 0.0538) – specifically, speakers with higher gayness ratings had lower f2 values. 4.5. vowel duration finally, vowel duration was examined, both as a mean across all stressed vowels analyzed for this project as well as individually within the stressed tokens of /æ/ and /ɛ/ discussed above. only one such comparison yielded statistically significant results, which was the duration of the second token of /ɛ/. in this case a one-way anova showed that the duration of this segment correlated significantly with speaker group (p < 0.0391) such that duration of this segment in the gsm group was significantly longer than it was for the ssm group. while the difference in duration was only statistically suggestive when comparing the gsm and tm groups, there was no statistical difference between the ssm and tm group, suggesting that tm are patterning more closely along the lines of the ssm 13 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 14 speakers rather than the gsm speakers. the effects of each group can be seen in table 5. table 5: speaker grouping vs. vowel duration (token #1 of /ɛ/) group comparison p-values ssm vs. gsm 0.0363 * ssm vs. tm 0.6755 . gsm vs. tm 0.0849 . . = suggestive at .1, * = significant at .05 these results serve as a final illustration of the ways in which the voices of tm and gsm are not alike acoustically. 4.6. summary this section has described several significant differences across the three speaker groups under investigation. first, for sibilant consonants, speakers in the tm group have a higher center of gravity for /s/ than the gsm and ssm groups, while speakers in the gsm group have a lower center of gravity for /ʃ/ than the tm and ssm groups. there may also be a connection between perceived masculinity and sibilant length and peak frequency. in terms of vowels, speakers in the gsm group had a significantly lower f2 for one instance of /æ/, but there were no differences in the two other examples. additionally, speakers with higher gayness ratings had higher f1 values for /ɛ/; it is also possible that speakers with higher gayness ratings had lower f2 values for this vowel. finally, speakers in the gsm group had a longer duration for the second token of /ɛ/ than speakers in either the ssm or tm groups. in terms of mean f0, f0 range, and overall mean f1 and f2, there were no significant differences across these groups. 5. discussion the results just presented point to a very significant conclusion that also serves to answer the second research question of this paper: given that we have established (in section 4.1) that the speakers in the tm and gsm groups are both perceived as gay-sounding compared to the speakers in the ssm group, are these two sets of speakers’ voices as similar acoustically as they are perceptually? 14 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 15 several findings in this study suggest that the answer is “no.” while some of the phonetic variables i have discussed correlated with the speaker’s mean gayness or masculinity rating regardless of group membership, suggesting that both tm and gsm speakers are using these features, other variables were used primarily by tm speakers or gsm speakers. specifically, tm speakers had a higher center of gravity for /s/, while the gsm speakers had a lower center of gravity for /ʃ/, a lower f2 for one instance of /æ/, and a longer duration for one token of /ɛ/. these results thus provide support for zwicky’s suggestion that there is more than one kind of gay-sounding voice, and that any number of different kinds of male voices that differ significantly from the straight-sounding norm can be perceived as indexing gay identity. the fact that trans men seem to be particularly likely to have gay-sounding voices even if they don’t identify as gay also suggests that early life gender socialization may very well be an important factor in accounting for why some men have gay-sounding voices and others do not – and, especially, why some men who do not identify as gay might nevertheless sound gay. of course, this isn’t to say that gender socialization is the only factor at work here. the fact that the trans men in this study did not identify as straight and for the most part tended to reject mainstream limitations on masculinity is surely relevant as well. however, considering that phonetic gender differences in the very features discussed here are known to arise early in life (e.g. flipsen et al. 1999 on /s/), socialization during childhood deserves more attention. at the same time, it isn’t as simple as saying that men who were raised as girls must be somehow inherently more feminine than men who were raised as boys – gender socialization does not have the same effect on everyone. if gender socialization always “worked” to produce gender normative adults, transsexuals would probably not exist, nor would many other kinds of gender diversity. why else would one trans speaker (#2 in appendix b) have such a high gayness rating at 4.167 out of 6, while another trans speaker (#4) had a much lower rating at 2.714? these two trans men are the same height, speak very similar varieties of american english, and are both queer-identified but have had little contact with communities of gay men. in fact, we might expect speaker #4 to have the higher gayness rating, because he was in a long-term relationship with a gay man at the time of recording, while speaker #2 has been in a long-term relationship with a straight woman for several years. clearly, these issues of gender, socialization, sexual orientation, and the interaction between them deserve more theoretical development than has so far been applied to this literature. unfortunately, a thorough exploration of these issues is outside the scope of the current project and will have to wait for future extensions of this work. there are also a few limitations of the study described in this paper that provide good reason for building on this work in the future. first, the small number of speaker subjects, especially in the gsm group, is an obvious weakness. in addition to more speakers, future work will also include more 15 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 16 features for acoustic analysis in order to achieve a fuller picture of how speakers talk. vowels in particular will reveal a more comprehensive picture of how these speakers vary from one another if examined as a system rather than in isolation. one final limitation is that read laboratory speech does a relatively poor job of reflecting how people actually speak in interaction. while smyth, jacobs and rogers (2003) found no significant difference between a spontaneous spoken passage and the dramatic passage used in this study, there is no reason to believe that even their spontaneous spoken passage produced in laboratory conditions would reflect a speaker’s more typical manners of speaking in everyday interaction. the findings put forth by podesva (2007) emphasize this point – he studied the use of falsetto as a stylistic resource for constructing a gay identity, but it is extremely unlikely that a speaker who makes use of falsetto during interaction with his friends, for example, would employ it while reading into a microphone, or even while engaged in a sociolinguistic interview. however, the results from this pilot will be highly valuable as a jumping off point for further work that makes use of interactional language data. one issue in this research that might be perceived as a problem is the fact that only one token of /æ/, out of the three analyzed, showed statistically significant variance across speakers and only one token of /ɛ/ showed significant differences in f2 and duration. however, this may simply be a reflection of normal intra-speaker variation – in other words, even the most gay-sounding speakers don’t necessarily sound equally gay all of the time. it may be that pronouncing a single word in a way that sounds gay is sufficient to create the perception that the speaker is gay.8 it is also worth noting that both of these tokens appeared in sentence-final position, while the other instances of these vowels were mid-clause. position in a syntactic or intonational phrase may thus play some part in determining which vowels are likely to be marked by this sort of sociolinguistic variation. extending the scope of analysis and including a greater number of speakers and tokens would also likely aid in answering these questions more satisfactorily. a final issue that may have complicated this analysis is the variability within the tm group. while these speakers were demographically similar (european-american, queer identified transsexual men in their early 20s of comparable physical size), the length of time since their transition varied considerably. one speaker (#2) began living in a male social role and taking testosterone approximately eight years before this recording was made (starting at age 15), another (#4) started testosterone approximately three years prior to being recorded (at age 19), and the last (#5) had started testosterone only 8 months prior (at age 20). mean f0 did correlate, among these speakers, with the length of time since they had started testosterone therapy. speaker #5, who began testosterone therapy at age 15, also had by far the lowest f0 among these speakers – indeed, 8 see mendoza-denton 2008 for an example of this phenomenon among latina gang members’ use of chicano english features. 16 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 17 the lowest of any speaker in this study – and this may be related to the fact that beginning testosterone therapy at an early age is thought to have more dramatic effects. however, these findings also suggest that previous studies on transsexual men’s voices (van borsel et al. 2000; adler & van borsel 2006) have been limited by only recording during the first year of testosterone therapy, as changes appear to continue beyond that point. to extend this work, i am currently completing the data collection phase of a larger project that builds on this pilot. in addition to a larger number of speaker and listener subjects, the analysis also includes a greater variety of acoustic measures, including voice quality, overall vowel expansion, a greater number of individual vowel classes with larger numbers of tokens for each class, and other sociophonetic features implicated in the linguistic construction of gender and sexuality. furthermore, a great deal more work is needed in order to explain, from a sociocultural linguistic perspective, why trans men tend to have gay-sounding voices and what this tells us about the indexical nature of gender and sexuality more generally. 6. conclusion in this paper, i have argued that the voices of trans men are perceived in much the same way as are the voices of gay-sounding non-trans men. however, there are important acoustic differences between these two groups in terms of both vowels and sibilant consonants. this supports a theory advanced by zwicky (1997) over a decade ago but which has yet to be fully integrated into scholarship on gay-sounding voices: there is more than one kind of gay-sounding phonetic style. further study is needed to confirm and expand on the results discussed in this paper, but the findings presented here are a promising starting ground for understanding the relationship between the many varieties of non-heteronormative voices. 7. references adler, richard k. and john van borsel. 2006. “female-to-male considerations.” in richard k. adler, sandy hirsch, and michelle mordaunt (eds.), voice and communication therapy for the transgender/transsexual client: a comprehensive clinical guide, 139-167. san diego: plural publishings. audacity development team. 2008. audacity (version 1.2.6.) [computer program]. http://audacity.sourceforge.net. boersma, paul and david weenink. 2008. praat: doing phonetics by computer (version 5.0.34) [computer program]. http://www.praat.org. 17 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 18 van borsel, john, griet de cuypere, robert rubens, and b. destaerke. 2000. “voice problems in female-to-male transsexuals.” international journal of language & communication disorders 35(3): 427-442. crist, sean. 1997. “duration of onset consonants in gay male stereotyped speech.” university of pennsylvania working papers in linguistics 4(3): 53-70. eckert, penelope. 2003. “language and gender in adolescence.” in miriam meyerhoff & janet holmes (eds.), handbook of language and gender, 381400. malden, ma: blackwell. fairbanks, grant. 1960. voice and articulation drillbook. new york: harper & row. flipsen, peter, jr., lawrence shriberg, gary weismer, heather karlsson and jane mcsweeny. 1999. acoustic characteristics of /s/ in adolescents. journal of speech, language, and hearing research 42(3): 663-677. gaudio, rudolf p. 1994. “sounding gay: properties in the speech of gay and straight men.” american speech 69(1): 30–57. gordon, bryan. 2008. “gay sounds: a non-discrete model of gay speech.” paper presented at the lavender languages and linguistics xv, washington, d.c., february 18. johnson, keith. 1997. acoustic & auditory phonetics. malden, ma: blackwell. levon, erez. 2006. “hearing ‘gay’: prosody, interpretation, and the affective judgments of men’s speech.” american speech 81(1): 56-78. levon, erez. 2007. “sexuality in context: variation and the sociolinguistic perception of identity.” language in society 36(4): 533-554. linville, sue ellen. 1998. “acoustic correlates of perceived versus actual sexual orientation in men’s speech.” folia phoniatrica et logopaedica 50(1): 35-48. mendoza-denton, norma. 2008. homegirls: language and cultural practice among latina youth gangs. malden, ma: blackwell. munson, benjamin, sarah v. jefferson, and elizabeth c. mcdonald. 2006. “the influence of perceived sexual orientation on fricative identification.” journal of the acoustical society of america 119(4): 2427-2437. munson, benjamin, elizabeth c. mcdonald, nancy l. deboe and aubrey r. white. 2006. “acoustic and perceptual bases of judgments of women and men's sexual orientation from read speech.” journal of phonetics 34(2): 202240. munson, benjamin. 2007. “the acoustic correlates of perceived masculinity, perceived femininity, and perceived sexual orientation.” language and speech 50(1): 125-142. pierrehumbert, janet b., tessa bent, benjamin munson, ann r. bradlow and j. michael bailey. 2004. “the influence of sexual orientation on vowel production.” journal of the acoustical society of america 116(4): 1905-1908. podesva, robert j. 2007. “phonation type as a stylistic variable: the use of falsetto in constructing a persona.” journal of sociolinguistics 11(4): 478-504. podesva, robert j., sara j. roberts, and kathryn campbell-kibler. 2001. “sharing resources and indexing meanings in the production of gay styles.” in 18 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 19 kathryn campbell-kibler, robert j. podesva, sarah j. roberts, and andrew wong (eds.), language and sexuality: contesting meaning in theory and practice, 175-189. stanford, ca: csli publications. r development core team. 2008. “r: a language and environment for statistical computing (version 2.7.2)” [computer program]. vienna, austria: r foundation for statistical computing. http://www.r-project.org. renn, peter. 2002. “subtypes of male homosexuality: speech, male sexual orientation, and childhood gender nonconformity.” unpublished ba thesis, university of texas at austin. sachs, jacqueline. 1975. “cues to the identification of sex in children's speech.” in barrie thorne and nancy henley (eds.), language and sex: difference and dominance, 152-171. newbury, ma: newbury house publishers. simpson, adrian p. 2009. “phonetic differences between male and female speech.” language and linguistics compass 3(2): 621-640. smyth, ron and henry rogers. 2002. “phonetics, gender, and sexual orientation.” proceedings of the annual meeting of the canadian linguistic association, 299-311. montreal, canada: l’universite du quebec au montreal. smyth, ron, greg jacobs and henry rogers. 2003. “male voices and perceived sexual orientation: an experimental and theoretical approach.” language in society 32(3): 329-350. waksler, rachelle. 2001. “pitch range and women's sexual orientation.” word 52(1): 69-77. zwicky, arnold. 1997. “two lavender issues for linguists.” in anna livia and kira hall (eds.), queerly phrased: language, gender, and sexuality, 21-34. new york: oxford university press. 19 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 20 appendix a the following is the passage from crist (1997) that was read by speakers. the italicized portion is the segment that was played for listeners and analyzed: you wouldn't believe what just happened! i was just sitting here studying, and it was getting pretty late, and i was going to go to bed here pretty soon. but then i started hearing these people screaming out in the street. so i got up, and i was going to yell out the window, "will you please hold it down out there!" but as soon as i poked my head out, i smelled smoke, and you know that ski store down at the end of the corner? it was all full of flames. there were all these people in the apartments upstairs screaming out of the windows; they must have been trapped. i was scared that the fire might spread down the street to my place too. then i heard sirens screaming, and all these cop cars and fire trucks pulled up. the firemen went up on ladders and helped all the people get out. one girl looked like she had bad burns on her skin, and this other guy fell, and the ambulance guys had to put a splint on his leg. i could see the guys down on the ground; they were having some kind of problem with the fire hydrant, but they finally got the hoses hooked up to the spouts, and then they went up and poked a hole in the roof with a big metal kind of stick, and they sprayed tons and tons of water in. it took them better than two hours to get the fire out. you know that spanish student down the hall from me? later, he told me he heard the owner set the fire himself. the whole thing was a big scam to get the insurance money. unbelievable! 20 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/3 doi: https://scholar.colorado.edu/cril/vol22/iss1/3/ female-to-male transsexuals and gay-sounding voices: a pilot study 21 appendix b the following table includes the mean ratings assigned to each speaker subject on a scale of straight to gay (1 = definitely straight, 5 = definitely gay) and masculine to feminine (1 = definitely masculine, 5 = definitely feminine). speaker # speaker group man gayness mean masculinity 2 tm 4.167 4.167 4 tm 2.714 2.857 5 tm 3.33 3.5 3 ssm 2 2 7 ssm 2 2.429 8 ssm 1.4 1.4 9 gsm 4.125 4 10 gsm 3.666 4 21 zimman: female-to-male transsexuals and gay-sounding voices published by cu scholar, 2010 colorado research in linguistics 6-2010 female-to-male transsexuals and gay-sounding voices: a pilot study lal zimman recommended citation microsoft word cril_zimman_revised.doc stories of narrative: on social scientific uses of narrative in multiple disciplines colorado research in linguistics. june 2007. vol. 20. boulder: university of colorado. © 2007 by john bryce merrill. stories of narrative: on social scientific uses of narrative in multiple disciplines john bryce merrill university of colorado this paper explores how narrative is understood and used by scholars in multiple disciplines to investigate social scientific issues. this is not, however, a traditional literature review. it is a report on an empirical study that involved systematic methods of data collection and analysis. the data in this case are scholarly literature on narrative, and an inductive analysis reveals three emergent themes. the first is the general tendency to view narrative as a formative mechanism in the construction of self and reality. the second addresses the ways narrative is conceptualized in terms of linguistic features, including structural and formal qualities, and how these features are studied in relation to social interaction. the third theme addresses how narrative is understood and employed as a method of social research. this paper contributes a valuable resource on narrative studies for scholars working within multiple disciplines. 1. introduction scholars of narrative understand that narratives are often both complex and revealing. they are linguistic structures: they are syntax and semantics; they are plots and characters; they are sequences. narratives are also substantive, in that they are what we say: they are phrases; they are colloquialisms; they are loaded. narratives, too, are contextualized within their construction: what they are depends on when and where they are said and, of course, by whom. narratives are ripe and fertile: they are simultaneously products of individual and society and individual and society are their products. narratives are social: they are local and national and global; they are feminine and masculine and all other positions possible. this laundry list of narrative’s qualities is not exhaustive—narratives are these things and many more—but even a list this brief implicates the limits of disciplinary narrative studies. it suggests that scholars interested in narrative must traverse disciplinary boundaries to do their work comprehensively. for example, we must consider simultaneously how sociolinguists theorize identity by studying linguistic practices; how anthropologists and sociologists speak to how local narratives i would like to thank leslie irvine, kira hall, martha gimez, and janet jacobs for their help with this paper. 1 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 2 resonate with or resist global ones; how social psychologists relate narrative to the construction and maintenance of selfhood; and how cultural studies implicates narrative in the discursive formation of salient social categories such as “heterosexual” or “poor.” if narrative is to be richly understood, its students must seek out and conduct research that crosses disciplinary boundaries. this paper provides a valuable resource for narrative scholars interested in crossing these boundaries. i explore how narrative is contemplated and used by scholars in multiple disciplines to investigate social scientific issues. using a systematic method of data collection and analysis, i focus on three predominant themes on the uses of narrative. the data in this case are scholarly literature on narrative, and an inductive analysis reveals these themes. the first is the general tendency to view narrative as a formative mechanism in the construction of self and reality. the second addresses ways narrative is conceptualized in terms of linguistic features, including structural and formal qualities, and how these features are studied in relation to social interaction. the third theme addresses how narrative is understood and employed as a method of social research. my ultimate aim is to encourage interdisciplinary studies of narrative by pointing to existing connections as evidence not only of the feasibility of this type of work, but its fruitfulness. while loosely united as social scientists, the authors i have referenced work in several different disciplines. these disciplines are characterized by varying theoretical and methodological assumptions. psychologists, for example, are generally interested in individual psychological processes, which may or may not be socially relevant or influenced, while sociologists place a primacy on society even when examining individuals. there are also substantive differences within disciplines. there are sociolinguists who pay little or no attention to social context when studying identity; others suggest it cannot be ignored. intradisciplinary difference is magnified in a field like social psychology, which requires qualifiers such as psychological social psychology and sociological social psychology to delineate critical even contradictory methodological and theoretical differences. a psychological social psychologist might run laboratory experiments to test theories of cognitive processes in simulated social settings, while a sociological social psychologist might study ethnographically the ways homeless people create meaningful relationships—two very different pursuits, both social psychological. variation in the social sciences is complicated further with the inclusion of newer fields like cultural studies, where disciplinary traditions do not formally exist and are often objects of cultural critique. altogether “social science” is a category that contains innumerable similarities, differences, and contradictions. and it is important to acknowledge, particularly when researching across these disciplines, that social scientists may have little more in common with each other than their shared title. for the purposes of this analysis, then, it is necessary to recognize that the authors i have referenced here, housed in different disciplines, are influenced by their larger disciplinary concerns. i have found that the differences in disciplinary 2 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 3 traditions—complicated, messy, and prohibitive in other ways as they are—have not prevented the emergence of a great deal of similarity in narrative studies across the social sciences. in fact, the heart of my argument is that at least in three important ways there is a great deal of consensus on the social scientific uses of narrative, regardless of disciplinary differences. therefore, i have avoided discussing disciplinary traditions when talking about a particular author’s work. instead, i have highlighted how authors’ works can be understood apart from their disciplinary moorings and as part of a larger, cohesive discourse on narrative. in other words i have given narrative center stage and have kept disciplines off in the wings. a final note before i proceed, the majority of the work reviewed is social psychological, but i hesitate to call it that because of the disciplinary implications. a clarification intended to stress subject matter, and not knowledge territories, lets me make a distinction between social psychology as the general study of individual and society and social psychology as a codified academic discipline. to be clear, my concern is with how social scientists address narrative’s place in the on-going relationships between the individual and society. the remainder of this essay is devoted to discussing research methods and analytical strategies employed in this project, presenting a summary presentation of the data and analysis, and concluding with closing thoughts. 2. methods this essay is more than a literature review. it is a report on data that is systematically collected and analyzed. the data in this case are literature on social scientific uses of narrative, and the analysis reveals existing interdisciplinary linkages in narrative studies. an objective of this essay is to present an important collection of narrative work to scholars who aspire to an interdisciplinary approach. in this way this project is a literature review. i also draw conclusions about narratives specifically and narrative research in general based on close analysis of the data. for this reason—the treatment of this project as an empirical case study—i am compelled to summarize my methods. my research here is mostly limited to the social sciences for two simple, yet complex reasons, which are practical and methodological limitations. narrative is so widely studied in the social sciences, and in its original home in the humanities, that exhaustive coverage is an unreasonable expectation. it would be impossible to review all that has been said about narrative given its enormous popularity. furthermore, all researchers either deliberately or indirectly exclude relevant data. ethnographers cannot talk with all groups of people that may shed light on similar meaningful practices. similarly, demographers cannot use all data sets to understand the ebbs and flows of migration. it is, perhaps, an implicit assumption in all research that some data are necessarily excluded. 3 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 4 the disciplinary literature analyzed here met an important initial criterion. my generic research interests are relationships between individual and society, subjects most commonly addressed in the social sciences. accordingly, i mostly limited my research to literature in the social sciences. specifically, i mined anthropology, cultural studies, psychology (including cognitive psychology), social psychology (including sociological and psychological forms), sociolinguistics, linguistic anthropology, and sociology for data on narrative. treatment of narrative in the humanities went largely unconsidered, except in such cases where literary scholars located their work in larger social scientific discussions. the most obvious absence is work on narrative as fiction. my search was even more narrowly focused on theoretical rather than empirical studies. this decision was guided by the want to understand what can roughly be referred to as the state of narrative studies across disciplines. i sought articles that summarized and synthesized narrative scholarship, often including references to empirical studies, offer a review of the treatment of narrative in specific fields. the empirical studies included here, such as penelope eckert’s (2000) ethnography of high school girls, offer rich overviews of narrative studies, often as introductions to their research. i also present case studies that exemplify theoretical ideas conveyed in this essay, though the focus remains narrative in general, even when specifically applied to case studies. having established these boundaries of selection, i employed two data collection strategies: theoretical and snowball sampling. these collection techniques are common among qualitative researchers, who are less likely than quantitative researchers to sample randomly. their popularity is in large part due to their potential to produce ample data. researchers use this approach when they have good theoretical reasons to search for data in particular places. guided by the aforementioned two key assumptions, i began reviewing literature in the usual fashion: searching social scientific databases, following bibliographic trails, and asking narrative scholars for their recommendations. formally the latter two methods of data collection are examples of snowball sampling, the practice of gathering data upon recommendations of others, usually research participants who are connected to potential participants. each of these practices yielded bountiful data, ultimately generating a data set consisting of forty-one journal articles, books, or book chapters. the data was analyzed using a strategy consistent with a grounded theory approach (charmaz and mitchell 2001; glaser and strauss 1967). this involved a recursive practice of data collection and analysis. data were sampled, reviewed, and initially loosely coded. i revisited and revised these categories as i collected more data. during this process, codes were assigned to emergent themes or, in other words, commonly held assumptions about narrative across selected disciplines. i ceased data collection when codes were solely recurrent instead of original. this is also a practice consistent with a grounded theory approach. i began with numerous codes that whittled the data down to six categories and ended with three master categories, which were produced by collapsing 4 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 5 smaller categories into larger ones. in essence, i created three dominant categories containing thematically related subcategories. initially, all varieties of narrative theories or propositions were documented as they appeared. for example, if a particular author discussed narrative time, then her work would be categorized as “narrative time.” eventually, commonalities among categories began to appear. for example, discussions of narrative time and linguistic variation were grouped under a large category, in this case the linguistically-oriented category forms and features. data were similarly coded and categorized until no new themes emerged. six themes emerged as recurrent and were labeled as follows: poststructuralism and structuralism; narrative construction of self; narrative construction of reality; narrative forms and features; narrative as interaction; narrative as method. based on additional theorized similarities, these were reduced to three categories: narrative construction of self and reality (ncsr); narrative forms and features (faf); and narrative as method (meth). these three “master” categories and their subcategories are the focus of the next section of this essay. 3. uses of narrative by analytical theme in this section, i present a detailed overview of the data and my analysis by discussing the data in thematic sections according to emergent themes. i begin with the largest section on narrative and the social construction of self and reality (ncsr), followed by a discussion of narrative as linguistic structures (faf), and concluding with narrative as a method of research (meth). in each section i outline the explicit meaning of the category and offer examples from the literature. i also present important discourses surrounding each theme, including commentary by proponents and opponents of these positions. i will not present in the body of the paper the arguments of every author analyzed; therefore, i have included a table (see appendix) that classifies authors by coded category. if an author or authors contribute to multiple categories, they are listed under each heading (e.g. riessman 1993 is located in all three master categories, so her name appears three times in separate columns). before covering narrative’s shared intellectual ground, let me speak to one of its most divisive, indecisive, and potentially pressing dilemmas: namely, arriving at an exact shared definition of narrative. as much work as has gone into defining narrative (for further discussions see bruner 1991; leiblich 1994; miller 1995; ochs and capps 2001) there has also been a great deal of disagreement: these disagreements are sometimes ideological and political (who gets to decide what is and what is not a narrative and what are the consequences of such decisions?); sometimes they are analytical (should narratives meet some strict criteria, such as possessing a beginning and end, notable events, cultural themes, and so on?); often, they are some confounding combination of each of these and more (if narratives must contain sequenced events, what about non-sequential talk told by people who do not or cannot tell sequential narratives, as in the chaos 5 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 6 narratives of terminally ill narrators in arthur frank’s [1995] the wounded storyteller). such an innocent question commands dizzying, sometimes insidious responses. but it or versions of it are asked repeatedly because it seems reasonable for researchers and theorists to want conceptual clarity. lacking clarity, ambiguous terms risk devalued explanatory power. thus, failure to reach an agreement on narrative’s definition could be the most pressing issue facing narrative scholars. or a consensus definition could be unimportant. this assertion may border on unfounded speculation (or treason to some), but it seems likely that one reason for this lacuna is that it is not vital to narrative studies to have a shared, concise definition. narrative studies are thriving without one, so clearly the explanatory value of narrative is not lacking. i think a better approach to this “problem” of narrative can be found by considering the question rather than the answer, specifically the type of question and the type of knowledge it is capable of producing. a definition of the term discourse is also hard to come by and contentiously contemplated. in a book devoted to defining critical terms in literary theory, paul bove (1995) writes an essay on why discourse should not be defined, essentially refusing the task at hand. the thrust of his argument is that discourse cannot be reduced to some meaningful essence. he begins justifying his contrary position by critiquing the question, taking the poststructuralist position that it comes out of existing “interpretive models of thought” that discourse studies seek to explore (53). in other words, one cannot ask innocently what something is, as i previously suggested. questions of this nature are born out of knowledge systems and power structures that dictate the limits of reasonable thought, of reason itself. it is only “reasonable” to ask what something is insofar as reasonable thinking falls within the boundaries of established modes of thought preserved in the power of institutions. it is reasonable to ask for a definitive version of discourse (or narrative) because contemporary thought values essential meanings (bove 1995, 53). what is the meaning of life? discourse studies are less interested in essential meanings; instead, they focus on “functional and regulative” (52) properties. for example, the question is not “what is discourse?” instead, we should ask, “what does it do?” or, as bove (54) suggests, what are its social and regulative effects? how does discourse function and how, as an analytical concept, does it discipline ways of thinking? this essay offers a similar way of thinking about narrative. it ignores the essentialist question “what is a narrative?” in favor of entertaining possible functions of narrative. it also does not address the epistemological dimensions of narrative studies, although this might be fertile ground for future research. instead, it concentrates on locating commonly held assumptions about what narrative does and can do. 6 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 7 3.1. narrative construction of self and reality (ncsr): the so-called “interpretive turn” in the social sciences has led scholars to question radically canonical ontological and psychological assumptions1. social reality is no longer assumed objective, and the notion of a ‘core’ self is also suspect. theoretical understandings of ‘reality’ and ‘self’ are pried from the hands of enlightened modern theorists and thrust into the spinning technicolor world of postmodernism, where, in the minds of some radical theorists, they are fictions or fossils. for the most part, however, scholars have opted not to annihilate these categories, in favor of deconstructing them to see what else can be learned about ‘reality’ and ‘self’. one of the most common and fruitful ways people have (re)envisioned self and reality is through the lens of narrative. narrative is not only seen as formative material for self and reality, but in some cases, a bridge between the two: between individual and society. the locus of the argument is that social reality exists because of human action, as do individual selves. communicative action is particularly critical. narrative as a form of communication, implicating what is said and how it is said in this process, then, is seen as being an essential conduit for the development of self and reality. the narrative construction of self and reality is not always addressed simultaneously, which was a reason for originally coding these two separately. so i will first review them separately, beginning with narrative and selfhood. next i will address narrative and social reality. third, i will add a section that qualifies the first two and adds to the overall theme by stressing each of these phenomena as types of interaction, narrative processes that must be enacted. the separation of these themes reflects my attempt to organize this section and not their empirical or theoretical differences. i will conclude this section by returning to the prevalence of these ideas in narrative studies and considering the few voices of dissent it faces. 3.1.1. narrative construction of self without reviewing the entire social history of the ‘self’ as a concept (see hewitt 1989), i want to point to a key development in the maturation of this concept, which is a generic shift away from social psychological notions of the self as a “core” entity, an object lodged psychologically or sociologically in the individual. modernist understandings of the self that sometimes figuratively, and sometimes literally, envision the self as an essence have been rejected by scholars 1 this is also sometimes referred to as the “discursive turn,” indicating a pointed focus on language. each references a marked move away from positivism, modernism, and objectivism, and an inclination to consider social realities differently. 7 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 8 looking to move away from essentialist social psychology and toward perspectives that stress the constructed nature of selfhood. this involves rejecting the idea that the self presupposes the social and thus social relations are guided by the internal drives of individual actors. the interiorized self (harre 1989) is a misleading fiction2. instead, selfhood is not betrothed to the individual; it is a social accomplishment. it requires the negotiated actions of individuals not only for its development, but also for its continued existence. there is no predominant theory of the constructed nature of selfhood, and there is disagreement among those who take this as canonical to social psychological studies. however, there is a great deal of consensus that narrative is a primary mechanism in the social construction and maintenance of self. holstein and gubrium (2000) suggest that selves are storied beings, the result of continued narrative practices that are undeniably social. bucholtz and hall (2005, 585), also arguing against organic “core” notions of selfhood, suggest that “identity [self] is the “product rather than the source of linguistic practices.” here bucholtz and hall (2005) rely on the concept of emergence to argue that selfhood, as well as culture and language, emerge during processes of interaction. it does not preexist interaction, but comes out of social performances. telling narratives—practices that rely on linguistic as well as relational skills—is one way selves come to be. 3.1.2. narrative as interaction it is critical to the proposition that selves are the products of narratives not to obscure the obvious point that narratives are products of narration, and that narration is a social activity. narratives cannot take on a reified quality, whereby they make us. they are creations, as much as we are. with an awareness of the performative nature of narrative self-construction, narrative scholars have paid considerable attention to unveiling how the telling of a narrative is just as important as the narrative produced. one of the more interesting developments to come out of this line of thinking is an interrogation of the putative differences between narrative and narration, or doing and saying. atkinson, coffey, and delamont (2003, 108), write that the “strict dualism between ‘what people do’ and ‘what people say’” held by researchers is at best unhelpful and at worst untrue. their point is that human actions are made understandable through narration; we tell stories of our actions to render meaningful what we have done. doing is saying. furthermore, narratives come into being by acts of telling. saying is doing. this second point emphasizes the interactive side of narrative 2 there are some who argue that the self in general is a fiction and no longer a salient social psychological concept. 8 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 9 and makes room for theorizing how narrative forms of action result in the creation of selves and social relations. holstein and gubrium (2000) outline a few generic strategies people employ to actively narrate identities. narrative linkage (108) is the first type. this involves weaving threads of coherence into stories that unify “biographical particulars” (108) and situational considerations3. narrators link present narratives to past and anticipated future ones, for example, to establish or maintain a consistent, desired self presentation. narrative slippage (109) references how narrators actively avoid (or employ) expected story lines or types of narrative. for example, narrative slippage occurs when individuals can claim the status “victim” but do not, when a story of victimization would be believable and accepted. instead, they may draw on qualitatively different discourses to narrate more (or less) favorable identities, such as “survivor”. this concept points to the agentive nature of narration, revealing how culture does provide means of narration, but individuals make decisions on what offerings they will use and for what reasons. finally, narrative options (110) describes how potential story lines are built into narratives, giving authors and audiences opportunities to accommodate the contingencies of narration. holstein and gubrium present an excellent empirical example of this concept in a narrative taken from an ethnographic interview of a student in a parent effectiveness class in a residential treatment center for emotionally disturbed children (110-2). the student, a mother, is asked whether she is like her parents in disciplining her children. her response supplies her a great deal of wiggle room: it depends. when my kids are really bad, i mean really bad, that’s when i think how my mother used to do with us. you know, don’t spare the rod or something like that in those days? but, usually, i feel that mother was too harsh with us and i think that kind of punishment isn’t good for kids today. better to talk about it and iron things out that way. still, like i say, it depends on how you want to think about it, doesn’t it? (from tanya quoted in holstein and gubrium 2000, 111-2). tanya leaves open the narrative option for either aligning herself or distancing herself from her mother. narrative options also speak to the agentic quality of narration and, like the previous strategies, this one reveals how narrative actions 3 holstein and gubrium’s concept is similar to jerome bruner’s (1990, 15) idea of “context sensitivity and negotiability,” which assumes that narratives must relate to the context in which they are told and should be negotiable. the difference here is that bruner uses his concept to define what a narrative should be; holstein and gubrium explicitly focus on the active creation of narratives, the focus of this section. 9 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 10 are contextual and contingent. thus, the narrative self is also contextual and contingent. these concepts exemplify how narration can be studied and understood as a process of self-construction. narratives told become storylines in the biography of self, one that is constantly under review and revision. this points to a triadic reflexive relationship with the self emerging between narration and narrative. of course, what type of self emerges and exactly how it does so is open for empirical as well as theoretical investigation. 3.1.3. narrative construction of reality bucholtz and hall (2004, 382) have developed a set concepts similar to holstein and gubrium’s, except their intention is to articulate how social relations are created during linguistic acts of self-construction. they refer to these acts as tactics of intersubjectivity (382). i want to present one of the three sets of tactics—adequation and distinction—to exemplify how social relations, including group memberships and communal identities, are the result of linguistic actions. adequation refers to “the pursuit of socially recognized sameness” (383). a blending of the words equation and adequacy, adequation requires narrating a reasonable likeness of others. in doing so, narrators must highlight available similarities while diminishing the significance of remarkable differences. adequation, then, refers to similarity among groups of people—nationalities, ethnicities, religions, and so on—and they are active creations rather than stable social categories. building generally on bourdieu’s analyses of the production and reproduction of class differences, bucholtz and hall (384) articulate distinction as “the mechanism whereby salient difference is produced.” similar to adequation, distinction involves selective punctuation of differences at the cost of recognizable similarities. thus, distinction is the active pursuit of difference even when evidence of similarity is available. using the concept of distinction, we can see how detrimental social differences that are often classified as inequalities are partially created and maintained as a result of narrative actions. bucholtz and hall theorize connections between linguistic strategies for identity construction and social relations constituted in part by these strategies. to put it another way, people tell stories to themselves and others and, in the telling, they create themselves and each other. they also create the very social realities in which they live. this is the narrative construction of reality. the proposition is that reality owes its existence in some or all part due to the narrative activities of people. it also assumes that narratives are ontological building blocks. in other words, reality is constituted by narration and consists of narratives. on the narrative construction of reality, it is necessary to make a distinction between moderate positions on reality construction and more radical ones. a moderate position on the narrative construction of reality, one that is more complimentary to theories that assume the existence of objective realities, is that narrative constitutes a type of reality. jerome bruner (1991, 4) proposes that 10 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 11 “narratives…are a version of reality” and are different from logical, scientific realities that are verifiable empirically. narrative realities, according to bruner, can achieve a likeness of reality, but do not exist in any verifiably objective way. only social validation determines the authenticity of narrative realities. not everyone who theorizes the ontological functions of narrative assumes a difference between narrated realties and objective ones. j. hillis miller (1995, 68) proposes an alternative way of thinking by presenting two possible versions of reality construction: one that suggests narrative creates reality, and one that argues it reveals it. the latter proposition implies the preexistence of a world that narrative can bring into focus. narrative translates blurry, incomprehensible realities into clear and meaningful ones. on the other hand, to suggest that narrative creates reality is to suspect the world does not presuppose narrative; narrative presupposes the world. the performative rather than clarifying function of narrative is reasonably considered a radical ontological view, one that stands in contrast to bruner’s theory of versions of reality and other theories that assume the existence of objective realities. regardless of disagreements over what types of realities owe their existence to narrative, there is a great deal of consensus that narrative and narrative activities produce consequential realties. 3.1.4. ncsr: popularity and dissent it is truly striking to consider how overwhelmingly common the sentiment is that narrative is essential to the formation of social reality, including the emergence and maintenance of self. what might be more remarkable than this is how few disagree with this proposition (see craib 2000, 64-74 for a scathing, but largely unconvincing critique). critics are less likely to engage in narrative studies directly, preferring to criticize the aforementioned interpretive turn in general. theories of narrative are but one part of a larger disagreement. interestingly, the most formidable and fruitful critiques have come from people wanting to present non-discursively oriented ontological and psychological theories. in these cases, the argument is not that narrative is not an important way that self and reality come to be, but that it is not the only way. nonetheless, the ontological and social psychological function of narrative is widely accepted and broadly used. if this is to continue, however, narrative researchers will have to consider seriously whether the role of narrative in the formation of the individual and society is overstated and, if other constructive processes are at work, how narrative can be seen in concert and/or opposition to them. 3.2. narrative features and forms (faf) consideration of the features and forms of narrative is at once focused on narrative structures and types of narratives and, simultaneously, on the nature of their existence. attention is paid to types of narratives: personal, local, cultural, canonical, and other forms. how narratives are composed and with what materials 11 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 12 are also important here. the grammar of narratives, paralinguistic qualities, and other formative features are investigated. unlike narrative studies that end with structures, research addressed in this section treats structure as suggestive of social psychological processes, such as identity work. a related, larger theoretical concern underscores much of this work, guided by the questions, “are narrative structures underlying entities, existing prior to human use or do they arise out of interactions?” how these questions are answered affects considerably how narrative structures are studied and what can be learned from them. i present this thematic section in three parts. first i discuss scholarship that argues narratives do contain universal, underlying structures, which can be found through structural analysis. i juxtapose this work with the work of those who assert that narrative structures do exist and are important objects of study, but that they are not essential things. narrative structures emerge during particular occasions of interaction—be they local or otherwise—and their existence depends on human action. i categorize the first position as a “structuralist” argument and the second as “poststructuralist”. i do recognize that there is more to these two categories of thought than what i am presenting here; however, i am only interested in their views on narratives structures. third, i look at work in this category that examines narrative features and forms, without regard for the nature of structures. 3.2.1. narrative and structuralism the thrust of a structuralist discussion on narrative is that certain indelible aspects of narrative, such as sequential order, morals, or plots, exist as universal structures. it is upon these structures that all narratives are built: they are essentially foundational. chatman (1978) refers to essential narrative components as “deep structure,” and “surface manifestation” occurs when stories are built upon them. variation in stories (or surface manifestations), even across cultures, is explained as mere differentiation, different spins of the same yarn. as mandlar (1984, 22) suggests, “stories have an underlying, or base, structure, that remains relatively invariant in spite of gross differences in content from story to story.” there are versions of shakespeare’s romeo and juliet told in different languages and times, by different people in different ways, but the core of the story does not change: regardless of the telling, it is still a tale of tragic destiny. it is important to consider that a structuralist argument envisions narrative structures existing at different levels of abstraction. deep structures are abstract analytical concepts, while surface manifestations (content) exist empirically. for example, william labov (1972) has famously argued that narratives are comprised of a series of clauses. a fully-formed narrative is comprised of an abstract, orientation, complicating action, evaluation, result or resolution, and a coda (363). these clauses are abstract categories that can take empirically different forms. an abstract may foreshadow death by poison, for example, or 12 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 13 good fortune from good deeds, but it will introduce audiences to forthcoming stories either way. the structuralist’s task is to reveal these structures through detailed narrative analysis, often involving parsing a narrative according to some criteria. for example, james gee (1986) proposes examining narrative data for the following structures: lines, stanzas, strophes, and sections. each of these units is defined in specific structural terms. for example, lines are short, simple clauses that typically begin with a conjunction and are syntactically and semantically related to lines around them (396). according to these specifics, a narrative is sectioned out into lines, as well as the other units. again, this method is an attempt to reveal analytically existing structural properties. and, to be clear, the analysis is only a means to a larger theoretical end. gee (1986, 2), like other structuralists, suggests that it makes sense that there is a great deal of cultural variation in the surface matter of stories.4 what also makes sense to gee is that there should be very little variation in the structure of these stories across languages. he states that [i]t seems hardly likely that there isn’t a great deal in common with the production of language in context across cultures, given that the same human brain, with its processing strengths and limitations, is producing this language in all cases (393). here is the heart of the structuralist argument: the human brain is the same in all people, and the human brain is the source of language and, thus narrative. therefore, the human brain must produce similar narratives for all people. if this is so, then these similarities can be found. their location is possible through structural analysis—in its many varieties—and so the location of a universal element of human cognition is similarly possible. narrative structures are cognitive structures, so cognitive structures can be revealed by looking at narrative structures. structuralists make claims about the universality of narrative structures and connect their existence to universal psychological processes. the general contentious issue here is whether these universal, underlying linguistic structures exist and, thus, can be located using structural analysis. moreover, if we accept the structuralist position on the existence of deep structures, we are compelled to consider their additional, more significant point: that these structures commonly 4 gee uses the term ‘discourse’ similarly to chatman’s ‘surface manifestations’. for the sake of consistency, i have stayed with chatman’s term or a version thereof. 13 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 14 bond all of humanity in an essential, organic way. human brains are made of the same matter, and so are our stories. 3.2.2. narrative and poststructuralism an informative way to transition from structuralism to poststructuralist narrative studies is to examine barbara hernstein smith’s (1981) treatment of the putative universality of the cinderella story. many structuralists (as well as other social scientists and literary scholars) have cited this as an actual example of a deep structure; rags to riches stories exist universally. smith, however, questions whether these stories share enough to merit common categorization and whether deep structures actually exist or whether they are actually the products of a particular type of knowledge, namely structuralism. her responses are decidedly anti-structuralism and lend themselves to poststructuralist thinking, although i am not sure whether smith would claim such classification. the deep structure of cinderella is the theme of “rags to riches.” this story can and has been told in a variety of ways. smith makes the wonderful, if not obvious argument that these variations often result in stories being markedly different. it is a stretch, she argues, to claim structural similarity when content changes so dramatically. she cites an icelandic “version” of cinderella, where the “prince” and “cinderella” invite the wicked stepmother to their ship for dinner; they serve her salted meat, which is the flesh of the wicked stepsisters that they just killed (1981, 212). this is certainly a grim version, but it is hardly comparable to the version of the brother’s grimm. it could still be a rags to riches story, but it could also be a story of the savagery of human nature. this is an interpretive decision that the analyst must make: it is not self-evident in the data. this second point—that analysts make interpretive decisions—is critical to smith’s position. not only do analysts make interpretive decisions, they do so within disciplinary boundaries. smith writes that [a]ll of us—critics, teachers and students of literature, and narratologists— tend to forget how relatively homogenous a group we are, how relatively limited and similar are our experiences of verbal art, and how relatively confined and similar are the conditions under which we pursue the study of literature (1981, 213). this is the lesson that feminists and others have passed on and that smith applies to structuralism: all knowledge is situated. theories of universal structures come from a particular group of people, structuralists, working in similar academic institutions and disciplines. thus, if a majority of literary scholars read all possible versions of cinderella and share the conclusion that they are structurally the same, one could assume this to be true. or one could assume that the theorized commonality of the stories more likely reflects the commonality of the theorists. rather than considering the intellectual merits of structuralism, then, it might be 14 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 15 more revealing to consider how this theory has been used, by whom, and for what reasons. whether smith identifies as a poststructuralist is not important. her critique of structuralism echoes foucault’s (1981) interrogation of the regulative functions of disciplines and, more importantly to narrative studies, his principles of specificity and exteriority (1981, 127). foucault, preeminent among poststructuralists, warns against considering the preexistence of discursive structures: [w]e must not resolve discourse into a play of pre-existing significations; we must not imagine that the world turns towards us a legible face which we would have only to decipher (127). this principle of specificity is buttressed by the principle of exteriority: that discursive explorations, including narrative studies, should not focus inwards to some mythical core of language; rather, they should remain externally concerned (foucault 1981, 127). this denial of interior structures of language and focus on what exists externally is what guides poststructuralist narrative research. poststructuralism is aptly named, as it moves beyond structuralism but retains some of its character. particularly, poststructuralist narrative scholars do examine narrative structures to study social psychological phenomenon, but they do so without heavy claims to universal cognitive processes. they discuss how structural qualities of narrative emerge during processes of interaction and how the uses and characteristics of these structures, such as how certain phrases are sequenced, are contextually dependent. groups may develop certain styles of narration that are marked by structural similarities, but they are the creators of their stories, not solely the creations of them. this departure from structuralism allows scholars to discuss how, as bucholtz and hall (2005, 585) propose, “identity is the product rather than the source of linguistic practices.” this position contradicts prior views of identity that suggested, for example, being a man encouraged speakers to tell masculine narratives. instead, telling “masculine” narratives is one way that people perform and become the category “man.” this also suggests that identities are not stable categories but malleable and relational social accomplishments. not surprisingly, similar thought exists surrounding discussions of the formation and maintenance of selfhood and other social realities, as referenced in the previous section. the guiding proposition is that identities and selves and other forms of social reality emerge in the processes of social relations, including narrative acts. not all research that avoids the essentialization of narrative claims to be poststructuralist. as well, not all poststructuralist researchers entertain questions regarding narrative structures. however, the debate between structuralists and poststructuralists is very important for narrative researchers, who are inevitably going to deal with structural questions. some may choose to move beyond these issues, but ignoring them is not likely or recommended. 15 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 16 3.2.3. narrative structures and types it is possible to consider narrative structures, as well as its other features and forms without considering their essence. for example, conversation analysis (see heritage 1984; holstein and gubrium 2000; sacks, schegloff, and jeffeson 1974) examines linguistic and paralinguistic aspects of narratives, such as narrative sequencing, turn-taking, and changes in vocal inflection, without assuming a structuralist or poststructuralist stance. however, these scholars have their own thoughts on the essential nature of narrative, which also do not go unchallenged. conversation analysts argue that it is at the microscopic level of ordinary talk that the mechanisms of social reality construction are found (holstein and gubrium 2000, 89). others have predictably refuted (or revised) these claims, suggesting that such a narrow focus excludes too much of social life to be so formative. regardless, the features and forms of narrative remain fruitful topics for social scientific investigation. concentration on narrative features might include sociolinguistic variation studies, where the researcher explores styles of speech, including prosody, lexicon, or syntax (eckert 2000, 1). penelope eckert explores ethnographically sociolinguistic variation among adolescent girls in a high school in new jersey. her theoretical aim is to bridge linguistic studies of structures with social studies of practice (44). she writes that variation is a linguistic process that is “inseparable from social process” (44). the “jocks” and “burnouts” of belton high narrate meaningful social realities by employing particular styles of narration. niko besnier (1992) also bridges the linguistic with the social in his examination of reported speech practices of nukulaelae, “a predominantly polynesian” group of people on a “small and isolated atoll of the tuvalu group” (164-5). reported speech, besnier argues, is often explored solely for its linguistic or grammatical qualities. besnier uses reported speech, the authorial practice of directly or indirectly quoting others, to explain how the nukulaelae satisfy the need to communicate affectively in spite of prohibitions against doing so. again, the focus is on how structural features of narratives are actively created and, most importantly, how these features reveal social processes. studies of narrative structures and social practices are plentiful. so, too, is research on forms of narratives. by “forms of narrative” i refer to identifiable types of narratives. these are sometimes divided into analytical binaries such as personal/cultural, everyday/dramatic, local/national. personal narratives can vary from those present during everyday conversations to ones given during life story interviews. cultural narratives reveal social meanings shared by a group of people. jerome bruner (1991, 19) theorizes a connection between personal and cultural narratives called “narrative accrual.” narrative accrual occurs when personal narratives amass into larger cultural narratives, taking on the qualities of collective sentiments. narrative forms defined geographically (e.g. local and national) are similar to the previous set. local narratives might be the shared stories of smaller 16 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 17 groups of people, even within nations, and national narratives are the stories of supposedly unified nations. national narratives may and often do contradict local narratives, but they must retain some resonance with local ones. this is why national narratives often manifest in generalities, potentially applicable and appealing to various groups. national narratives, often products of mass media, attempt to maintain a hegemonic dominance over local narratives (jacobs 2004), because their function is to maintain existing social relations, which are in large part narratively formed. ronald jacobs’ (2004) study of narrative and public culture references the 1992 uprising in los angeles to show how national narratives shaped local understandings. particularly, national narratives colored the incident as either a localized problem of chaotic violence or the expected outcome of the rodney king trial, where one racist cop, mark furman, or a racist jury could be blamed. competition from non-national narrators who might have attributed the incident to institutionalized racism and poverty was rendered largely impotent. narrative forms are also referred to as dramatic (textual) or everyday (conversational). elinor ochs and lisa capps (2001) propose that these types of narratives are different in three important ways: process of construction, prevalence, and ontological function. unlike dramatic narratives that are thought to be systematically and intentionally constructed, everyday narratives take on more chaotic qualities. they are often collaboratively produced in unscripted instances of interaction, with authors changing positions with audiences sometimes unexpectedly. the messiness of everyday narration offends the sterility of dramatic narrative construction. everyday narratives, according to ochs and capps (2001, 3) are far more ubiquitous than dramatic ones, marking a clear difference in the prevalence of the two. finally, a qualification combining the first two, the hazards of everyday narratives and their abundance suggest that they play a more dominant role in sense-making activities. therefore, everyday narratives are a primary ingredient in the making of social realities; dramatic, scripted, rehearsed, controlled narratives offer secondary contributions. whether these distinctions—or any distinctions—between everyday and dramatic narratives hold up is questionable. however, their differences are typically met with few objections from scholars or general audiences. i conclude this section by demonstrating a connection between each section in this thematic category “narrative as structure.” with or without a theory of the existence of narrative structures that presuppose social relations, narrative researchers have richly explored, as barbara johnstone (1990, 77) puts it, “how storytellers make use of the resources of grammar to make statements about, and to manipulate, social relationships in their stories and in the world” (77). if we expand johnstone’s “resources of grammar” to include additional narrative resources (linguistic structures as well as types of narratives), we can see the type of recursive relationship between individual and society that social psychologists strive to understand. individuals create narratives in particular ways by drawing on resources, such as existing cultural narratives. likewise, the cultural narratives 17 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 18 of any society have their genesis in personal narratives, taking on forms and meanings created by individuals. the assumption here is that individual and society are in an on-going dialogic relationship. dennis tedlock and bruce mannheim argue that [c]ultures are continuously produced, reproduced, and revised in dialogues among their members. cultural events are not the sum of the actions of their individual participants, each of whom imperfectly expresses a pre-existent pattern, but are scenes where shared culture emerges from interaction (1995, 2). what these authors refer to as the “dialogic emergence of culture” is synonymous with my claim that individual and society exist in a recursive relationship whereby the narrative construction of reality occurs in part at the level of structural narrative usage. one important implication of this assumption is obvious: narrative studies are critical to understanding the emergence and continued existence of social life. exactly how to use narrative to answer this core social scientific question is not so obvious. 3.3. narrative as method (meth) in this section, i consider researchers’ uses of narrative methods to collect data, how narrative data is analyzed, and ancillary methodological considerations. some of the items discussed here will be relevant to other methods of research, particularly qualitative methods. however, this section addresses discussions of research that refer explicitly to narrative as a type of method. this is consistent with my desire to represent as genuinely as possible the data on narrative. of course, as riessman (1993) and others point out, honestly representing narrative data is hardly a simple task. this thematic section can be separated into three smaller categories: 1) narrative as a method of data collection, 2) narrative as a method of analysis, and 3) methodological issues in doing narrative research. these can be separated for purposes of summarizing, but these matters are closely related. narrative methods are used to generate narrative data that can be subsequently analyzed in a particular way. guiding and sometimes inhibiting these processes of collection and analysis are ethical and methodological issues that all narrative researchers are likely to encounter. so, i’ll treat these categories separately to begin with but conclude with thoughts on their interrelations and the implications for narrative studies in general. narrative as a method of data collection is best exemplified by william labov’s groundbreaking work (1972). labov devised a method for collecting narrative data that involved asking participants a leading question, such as “when was a time where you nearly experienced death?” labov, however, was 18 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 19 uninterested in the empirical intricacies of near-death experiences. these questions were mere means for extracting narratives, which would be analyzed not for their content but their structural features. the key feature of this approach is that narrative content (what is said) is a secondary concern to form (how it is said). this approach is dissimilar to interviewing methods where the object of exploration is the substantive details of participants’ narratives. instead, as riessman (1993,2) argues, this method produces data that speaks to how narratives are composed, what narrative resources are used, and how narrators convince audiences of authenticity. gubrium and holstein (1998 cited in holstein and gubrium 2000, 104) refer to these as dimensions of “narrative practice.” they write that this term characterizes “the activities of storytelling, the resources used to tell stories, and the auspices under which stories are told.” narrative method, then, can be considered a method of observing narrative practices. narrative analysis involves the ways researchers draw theoretical conclusions from narrative data or, in other words, how specific narrative practices are conceptualized. leslie irvine’s (1999) narrative study of codependents anonymous groups reveals how group members use the vocabulary of the group to construct a “codependent” self. loseke (2001) also writes about how “battered women” sometimes draw on cultural narratives (formula stories, in loseke’s term) to tell an acceptable story of victimization, which is needed to secure services in domestic violence shelters. in each case, the narrative practices of individuals are revealed and conceptualized theoretically. loseke and irvine’s analyses of narratives reveal how self and identity are accomplished using narrative. riessman (1993, 13) proposes that narrative analysis is one of five stages of narrative research. in fact, it is just one stage in the process of representing narrative experiences. she proposes five stages (or types) of representation: 1) attending, 2) telling, 3) transcribing, 4) analyzing, and 5) reading. this methodological assertion begins by assuming that researchers attend to experience selectively. we cannot make sense of everything around us, so we make sense of some things. our choices largely reflect who we are, including social positions we occupy (gender, sexuality, age, and so on). experiences are then told to others, a process that is also infused with subjectivity—ours and our audience’s. the character of narratives depends on who is listening (or reading), as much as who is telling. researchers often transcribe the telling of experiences, and it is commonly assumed that the act of transcription is unproblematic. voices are turned into words. but, as bucholtz (2000, 1463) demonstrates, “the transcription of a text always involves the inscription of a context.” transcribing requires interpretive decisions, from deciding how narratives will be transcribed (with or without temporal indicators? with or without notations for changes in vocality?) to what will be transcribed (will the whole narrative be transcribed or just parts? will utterances be included?). riessman and bucholtz’s point is that transcribing is no neater, no less objective than the other levels of representation. of course, neither is conducting a narrative analysis. to make matters more complicated, 19 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 20 audiences that read narrative research make up riessman’s (1993, 14) fifth level of representation, thereby adding additional interpretive contingencies to the subjective mix. all of this leads to the conclusion that narrative researchers must be aware that their tasks are inherently charged with subjectivity and lodged in particular social relations. they are also charged with being less than scientific, an offense that sometimes threatens exile from the academy. narrative research is often confronted with claims questioning its validity as a method of social science research. “given that much of it moves beyond the realms of realism and positivism,” social scientists ask, “what criteria exist to judge credible narrative work?” riessman (1993, 65-68) proposes these possible criteria. the first is persuasiveness. the question asked to audiences that included research participants and like scholars is this: are the data and analysis persuasive? participants may judge how their voices are represented, empirically and analytically. other scholars can consider how the research fits in with other similar literature. if both parties are persuaded, then one criterion for validity is met. next, akin to persuasiveness is correspondence. do theories derived match the data? again, this question should be asked of participants and other researchers. finally, the work may be judged valid if it can be useful to future research. this usefulness is determined by related researchers who, presumably, would consider the previous standards of validity. riessman’s proposal, although not entirely unique to narrative research, does provide a sold initial stance for defending against accusations from social scientists that narrative should, figuratively speaking, go back where it belongs— in the arts, not the sciences. it also rightly avoids one of the least convincing complaints about narrative research: that people lie. ian craib (2004) uses the academic euphemism “bad faith narratives” to shroud his complaint about lying in sophisticated language. i admit to finding his tongue-in-cheek comment that his “mother may have been a better psychologist than [jerome] bruner for she could tell the difference between a life lived and a life as told” (65) humorous. however, the sentiment—that a true reality exists and it is experiential—is not as welcomed. the fatal flaw in this argument is that it fails to leave the confines of positivism and realism to critique narrative research on its own terms. criticisms of this kind do not advance narrative research; they only undermine it. at best, they allow narrative a place in the softer side of academia. i am not suggesting that theories of narrative go unquestioned by outsiders. in fact, i think it is vital for both narrative and non-narrative scholars to interrogate theories of narrative, such as narrative’s relationship to the construction of reality. our methods of research and our analytical techniques should be continually scrutinized and, if necessary, revised. my point is that critiques should be constructive; they should be guided by the objective of advancing narrative studies and, consequently, enriching social scientific knowledge. 20 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 21 returning to craib’s concern about lying, a valid methodological question would be, “how do narrative researchers consider the truth or falsity of narrative data?” richard bauman (1996) handles this question in a way that results in a work of innovative narrative research. bauman argues that the issue of truth is presented as a typological problem: one type of narrative is the truth; lies are another type. instead, bauman proposes that the question is ethnographic. “what is needed,” bauman (1996, 161) writes, “are closely focused ethnographic investigations of how truth and lying operate as locally salient storytelling criteria within specific institutional and situational contexts in particular societies.” this is exactly what he does with his study of expressive lying among dog traders— lying, like telling the truth, is one way narrative lives are lived. this type of response should be the archetype for constructive reactions to legitimate critiques. 4. conclusion this essay points to studies that implicate narrative in the formation of reality and in the creation and maintenance of selfhood. it also summarizes how narrative as linguistic structures and forms are used by individuals to create meaningful social relations. finally it addresses how empirical and theoretical knowledge of narrative is generated and how this knowledge can be valued in the social sciences. it does not come close to clarifying narrative’s definitive character, and may in fact make the question “what is narrative?” even harder to answer. hopefully, it discourages the question altogether, in favor of inquiring into the social function of narrative. only a few answers to this question have been presented here. so many more answers—some contradictory, some complementary—are to be found both within disciplines and between them. with this essay, i have hopefully provided a helpful resource for students of narrative who prefer to cross disciplinary boundaries rather than stay within their own territories. i have done this by providing a synthesis of narrative scholarship and references for additional research. if my analyses and summaries are believable, researchers have an invaluable tool for future investigations. if they are not, the data is available for alternative considerations. references atkinson, paul, amanda coffey, sara delamont. 2003. key themes in qualitative research. walnut creek, ca: alta mira. besnier, niko. 1992. "reported speech among the nukulaelae atoll." in jane h. hill and judith irvine (eds.) responsibility and evidence in oral discourse, 161-181. cambridge, uk: cambridge university press. bove, paul. 1995. "discourse." in frank lentricchia and thomas mclaughlin (eds.) critical terms for literary study, 50-65. chicago: university of chicago press. 21 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 22 briggs, charles l. and richard bauman. 1992. "genre, intertextuality, and social power." bruner, jerome. 1991. "narrative construction of reality." critical inquiry autumn:1-21. bucholtz, mary. 2000. "the politics of transcription." journal of pragmatics 32:1439-1465. bucholtz, mary and kira hall. 2004. "language and identity." in alessandro duranti (ed.) a companion to linguistic anthropology. malden, ma: blackwell. —. 2005. "identity and interaction: a sociocultural linguistic approach." discourse 7(4-5): 585-614. charmaz, kathy and richard g. mitchell. 2001. "grounded theory in ethnography." in amanda coffey, paul atkinson, sara delamont, john lofland and lyn lofland (eds.) handbook of contemporary ethnography, 160-74. london: sage. chatman, seymour. 1978. story and discourse: narrative structure in fiction and film. ithaca, ny: cornell university. craib, ian. 2000. "narrative as bad faith." in shelly day sclater, molly andrews, corrine squire, and amal treacher (eds.) the uses of narrative: explorations in sociology, psychology, and cultural studies, 64-74. new brunswick: routledge. eckert, penelope. 2000. linguistic variation as social practice. malden, ma: blackwell publishers. foucault, michele. 1981. "the order of discourse." in robert young, untying the text, 48-78. boston and london: routledge and kegan paul. frank, authur w. 1995. the wounded storyteller. chicago: university of chicago press. gee, james paul. 1986. "units in the production of narrative discourse." discourse processes:391-422. glaser, barney g. and anselm l. straus. 1967. the discovery of grounded theory: strategies for qualitative research. chicago: aldine. gubrium, jaber f. and james a. holstein. 1998. "narrative practice and the coherence of personal stories." sociological quarterly:163-87. harre, rom. 1989. "language games and texts of identity." in john shotter and kenneth gergen (eds.) texts of identity, 20-35. london: sage. heritage, john. 1984. garfinkle and ethnomethodology. cambridge, england: polity. hewitt, john p. 1989. dilemmas of the american self. philadelphia: temple. holstein, james a. and jaber f. gubrium. 2000. the self we live by: narrative identity in a postmodern world. new york: oxford university press. jacobs, ronald n. 2000. "narrative civil society and public culture." in shelly day sclater, molly andrews, corrine squire, and amal treacher (eds.) the uses of narrative: explorations in sociology, psychology, and cultural studies, 18-35. new brunswick: routledge. 22 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 23 johnstone, barbara. 1990. stories, community, and place: narratives from middle america. bloomington: indiana university press. labov, william. 1972. lanaguage in the inner city. phildadelphia, pa: university of pennsylvania. lieblich, amy. 1994. "introduction." in ruthellen josselson and amy lieblich (eds.) exploring identity and gender: narrative study of lives, xi-xiv. thousand oaks, ca: sage. loseke, donileen r. 2001. "lived realities and formula stories of 'battered women'." in jaber gubrium and james holstein (eds.) institutional selves: troubled identities in a postmodern world. new york: oxford university press. maines, david. 2001. faultline of consciousness: a view of interactionism in sociology. new york: aline de gruyter. malson, helen. 2000. "fictional(ising) identity? ontological assumptions and methodological productions of ('anorexic') subjectivities." in shelly day sclater, molly andrews, corrine squire, and amal treacher (eds.) the uses of narrative: explorations in sociology, psychology, and cultural studies, 150-163. new brunswick: routledge. mancuso, james l. 1986. "the acquisition and use of narrative grammar structure." in theodore r. sarbin (ed.) narrative psychology: the storied nature of human conduct, 91-110. new york: praeger. mandlar, j.m. 1984. scripts, stories, and scenes: aspects of schema theory. hillsdale, nj: lawrence earlbaum associates. mannheim, bruce and dennis tedlock. 1995. "introduction." in bruce mannheim and dennis tedlock (eds.) the dialogic emergence of culture. urbana and chicago: university of illinois press. miller, hillis j. 1995. "narrative." in frank lentricchia and thomas mclaughlin (eds.) critical terms for literary study, 66-79. chicago: university of chicago press. neisser, ulric. 1994. "self narratives: true and false." in ulric neiser and robyn fivush (eds.) the remembering self: construction and accuracy in the selfnarrative, 1-18. cambridge: cambridge university press. ochs, elinor and lisa capps. 2001. "a dimensional approach to narrative." in living narrative: creating lives in everyday storytelling. cambridge, ma: harvard university press. polkinghorne, donald e. 1991. "narrative and the self-concept." jounral of narrative and life history:135-53. riessman, catherine kohler. 1993. narrative analysis. london: sage. sacks, harvey, emanuel schegloff, and gail jefferson. 1974. "a simplest systematics for the organization of turn-taking in convesation." language 50:696-735. sampson, edward e. 1989. "the deconstruction of the self." in john shotter and kenneth gergen (eds.) texts of identity, 1-19. london: sage. 23 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 24 sarbine, theodore r. 1986. "the narrative as root metaphor for psychology." in theodore r. sarbin (ed.) narrative psychology: the storied nature of human conduct, 3-21. new york: praeger. schaefer, roy. 1981. "narration in the psychoanalytic dialogue." in w.j.t. mitchell (ed.) on narrative. chicago: university of chicago press. smith, hernstein barbara. 1981. "narrative versions, narrative theories." in w.j.t. mitchell (ed.) on narrative, 209-232. chicago: university of chicago press. somers, margaret r. 1994. "the narrative construction of identity: a relational approach." theory and society 23:605-49.:605-49. 24 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/4 doi: https://doi.org/10.25810/cnf0-ks61 stories of narratives 25 appendix authors by master code faf ncsr meth briggs and bauman 1992 atkinson, coffey, and delamont 2003 atkinson, coffey, and delamont 2003 bruner 1991 bruner 1991 bucholtz and hall 2004 bucholtz and hall 2004 bauman 1996 bucholtz and hall 2005 bucholtz and hall 2005 bucholtz 2000 chatman 1978 foucault 1981 graves gee 1986 harre 1989 gubrium and holstein 1998 foucault 1981 holstein and gubrium 2000 holstein and gubrium 2000 harre 1989 johnstone 1990 holstein and gubrium 2000 maines 2001 jacobs 2000 johnstone 1990 mancuso 1986 labov 1972 jacobs 2000 miller 1995 lieblich 1994 labov 1972 neiser 1994 maines 2001 heritage 1984 oaks and capps 2001 irvine 1999 maines 2001 polkinghorne 1991 loseke 2001 mancuso 1986 riessman 1993 neiser 1994 mandlar 1984 sampson 1989 oaks and capps 2001 mannheim and tedlock 1995 sarbin 1986 riessman 1993 sacks, schegloff, and jackson 1974 schafer 1981 sampson 1989 besnier 1992 ochs and capps 2001 schafer 1981 polkinghorne 1991 riessman 1993 loseke 2001 smith 1981 25 merrill: stories of narrative published by cu scholar, 2007 colorado research in linguistics 6-2007 stories of narrative: on social scientific uses of narrative in multiple disciplines john b. merrill recommended citation untitled microsoft word ryan-cril2021-proof_final.docx 1 what led us to the data zachary ryan university of colorado boulder the usefulness of bibles in natural language processing is wildly underrated, especially when looking at low-resource languages. the reason these texts are so useful in machine translation is because it provides parallel text in languages that would otherwise have non or very little other parallel texts available. the question is then, how can we turn these parallel texts into a useable, organized, and reliable data set? there are two general methods that can be used to collect and organize this data, the old fashioned way by hand or automating the process through a computer. while this paper will touch on both methods, it takes a much deeper dive in the automated processing side and looks at some of the reasons this route was chosen. keywords: data collection, data organization, machine learning, natural language processing, automation 1. introduction many researchers and companies have ideas for machine learning or artificial intelligence (ai) projects which often times hit a barrier immediately. where do i get enough data to create my model? this is the very question myself and two other researchers had to ask ourselves when working with what we consider a low-resource language. for those not familiar with the term low-resource language, it is a language that lacks significant monolingual or parallel corpora that usually requires manually crafted linguistic resources needed for the study of the language. the languages the lead of my team chose to study were basque and navajo, which we wanted to use to create working translation models to english from the respective language and vice versa. 1.1. the team and its research the three of us started working together to because the leader of our team, mans hulden, a linguistics professor at university of colorado, wanted to see how useful bibles could be for neural machine translation (nmt) models. and if any tokenization methods on top of this could help improve the nmt models. the second member of the team, ling liu, a master’s student at university of colorado, was already helping mans with the linguistic side of the project before i was brought in. i was the last person brought in to help more with the coding side of the project colorado research in linguistics, volume 25 (2021) 2 because i had done some previous work in natural language processing (nlp) with mans and was also at the time a senior finishing my computer science degree. 2. choosing a method for creating translation models to create this type of translation model would be possible through two methods. large amounts of monolingual text which can be looked at for each language where patterns are found until small accurate translations are made which are then expounded upon. this will eventually lead to very accurate translation models but this has only been shown to be successful with languages with large amounts of text like english and spanish (conneau et al. 2018). the second way to create a successful translation model is to use sets of parallel text from each language. since the translation is known, the model guesses and we tell it if its right or wrong and to adjust from there. this method does not require as nearly as much text as the first option. since we were working with low-recourse languages we chose the second option to create our translation models. 3. finding data the next steps were to now find where we could collect data from. one of us had found a pdf file of pictures of bible scripture written in navajo. an example of what one of these pages looks like can be seen below in figure 1. figure 1 what led us to the data 3 we attempted to use an optical character recognition (ocr) reader on the pdf document to pull the text from it. what it did extract from the document was messy data with missing characters and also extra characters being picked up from the poor quality of the photos. an example of a messy page and its incorrect results from the ocr reader can be seen in figures 2 and 3 respectively. one of the mistakes can be found in the last word of the second line in figure 2 and 3, can you find any more on this page? figure 2 figure 3 this would not be useful unless someone manually went through the text to fix the mistakes the ocr reader made. another method would need to be found, but this gave us the idea to search out more bible scripture for a couple of reasons. the bible has been translated to many languages and even languages only spoken by very remote people in hopes of bringing those people to the religion. the religious aspect is not what is important here but the fact this would provide parallel translations for many languages that had very little other forms of written text. we also had a very colorado research in linguistics, volume 25 (2021) 2 high confidence that this data would be accurate because it is very likely the church would not want their message to be misconstrued. the last thing that enticed us was the bible and its translations are all public, so there is no worry of using someone's private data or data you may have to pay for. this led us to a website, www.bible.com, that contained thousands of translations for the bible, some of which were considered low-resource languages. basque and navajo both had translations for sections of the bible which we wanted to use; the issue was how do we collect it all. 3.1. data collection and processing there were two ways that i saw we could collect this data: manually copy and pasting the data to a text document which could then be processed or create a web scraper for this website. with my two colleagues being linguistics researchers and myself being a computer scientist i was tasked with collecting the data. i chose to use the web scraper and will go though some of the reasons why and the code in this section, the data storage, and additional features sections. i wanted to use the web scraper because in the long run it would be easier than manually coping the text and would also allow for some preprocessing to be done alongside data collection. due to the structure of the website, it allowed for a somewhat easy automated data collection process, not that this cannot be done with other websites, but i'll explain what i mean. the website we used defines which bible we used through a numeric code, the biblical book, and the language version used all in the url of the page. this allows us to go directly to the bible version and chapter we want without needing to go through intermediary pages. with this initial part figured out this meant there were two next steps. how do we get it to move from to the next chapter of the bible version wanted and what do we actually scrape from the page? taking on the problem of moving to the next chapter i anticipated would be more difficult so i started with this one first. after looking at the bible versions for the two respective languages i noticed that not every chapter of the bible was translated for the respective languages. to work around this, i obtained a list of the chapters and the abbreviation used in the url embedded in the html of the website. i could then create an array of the abbreviations to use to ping the website in the web scraper. the outermost layers of the web scraper consist of loops used to ping the website by cycling through the list of abbreviations to complete the url. if a ping was successful this indicates that the bible version contained the chapter pinged and the html of the page would be grabbed, this process will be talked about later. if a ping was not successful this would mean what led us to the data 3 the chapter is not there and to move onto the next. at this point in the development process the only other feature added to this layer of the code was being able to choose specific ranges of chapters to ping and possibly grab. we wanted to start collecting data in any format so i moved onto the next problem. 3.2. data storage the next major problem was what did we want off of the page and how should the data be stored. the way the data was obtained was through a python package called beautifulsoup which would grab the entirety of the html for the page pinged. after reading through the html, i created a regular expression that would pick out the text of the titles and the verses based off of the html tags the text would reside in. after getting to this point, i needed a way to store the data. the website would sometimes provide the title of the chapter on the page and also sub-titles depending on the version of the bible. another categorizing feature given was the verse numbers of each chapter. using all of these categorizing features in conjunction with one another provides a simple way to catalog the data and also this would provide enough break down of the text so that it would give an ample amount of data points but also each data point would have some depth to it. the storing of the data was done through text files which consisted of two columns. the first columns were numeric reference consisting of the book version, chapter number, verse number, and an indication if it was a title or the verse itself, these values were separated by a colon. the second column was the cleaned text of the verse. here the text needed to be cleaned of irregular characters, these were things like commas and dashes that were changed to comply with the machine learning tools being used. at this point the web scraper was working in its most primitive form. it was capable of connecting to the version of the bible needed if you knew the numeric code, could then collect all the information for that specific bible, and store the information into a more user-friendly version. 3.3. additional features of data collection and processing after refining the code somewhat and talking with the other researchers we wanted to add more features to the web scraper. the two major tasks were to make the web scraper more universal so that it could be useable on any of the bible's available and to also have the preprocessing work be done alongside the initial storing of the data. at this point i opted to give the file command line options to make the web scraping process more user friendly. i added features to scrape all colorado research in linguistics, volume 25 (2021) 4 available bibles, create specific lists of bibles to grab, range of bibles to grab, an option to tokenize words based off english grammar, an option to tokenize based of a standard rule set, and a syllabifier. a more hidden feature available if you want to write some of your own code for a language was to create your own tokenizer. during the preprocessing of data within the web scraper another file gets called to tokenize the text. within this file tokenizers can be added for specific languages and if one is not present the standard rule set is used unless instructed otherwise. we experimented with a few different types of tokenizers, all of which can still be seen in the original code. at this stage the code was much closer to its final form, from here only bugs and small formatting changes in file structure used to save the text were changed in the code. from here there was nothing left to do but collect our data and begin building our models. 4. conclusion now that we were able to collect and preprocess our data, we could actually run the experiments that mans and ling had originally set forth. is it possible to create neural machine translation model from a low-resource language and can any additional techniques be applied to help aid this? the results of our research can be found in our paper (liu et al. 2021) and the code itself can be seen on our github (liu et al. 2021). references conneau, alexis; guillaume lample; marc’aurelio ranzato; ludovic denoyer; and hervé jégou. word translation without parallel data. arxiv:1710.04087 [cs.cl] (january 30, 2018). accessed april, 2021. online: http://arxiv.org/abs/1710.04087. read the bible. a free bible on your phone, tablet, and computer. read the bible. a free bible on your phone, tablet, and computer. | the bible app | bible.com. (n.d.). https://www.bible.com/. liu, l., ryan, z., & hulden, m. (2021). the usefulness of bibles in low-resource machine translation. proceedings of the workshop on computational methods for endangered languages, 1, 44–50. https://doi.org/10.33011/computel.v1i.957 liu, l., ryan, z., & hulden, m. (2021). the usefulness of bibles in low-resource machine translation. github repository, https://github.com/lonelyrider-cs/low_resource_mt rhematization as etiology in the diagnosis of posttraumatic stress disorder rhematization as etiology in the diagnosis of posttraumatic stress disorder cover page footnote this paper was only possible through the generous feedback and reference suggestions by kira hall, kathryn goldfarb, and chase raymond. this working paper is available in colorado research in linguistics: https://scholar.colorado.edu/cril/vol24/iss1/6 https://scholar.colorado.edu/cril/vol24/iss1/6?utm_source=scholar.colorado.edu%2fcril%2fvol24%2fiss1%2f6&utm_medium=pdf&utm_campaign=pdfcoverpages rhematization as etiology in the diagnosis of posttraumatic stress disorder ayden parish university of colorado boulder current psychiatric nosology emphasizes observable symptoms as the central schema by which mental illnesses should be classified; patients are identified as depressed or schizophrenic by virtue of observed behavior or reported experiences, rather than theoretical underlying causes that may lead to an array of diverse presentations. within this schema, the diagnosis of posttraumatic stress disorder (ptsd) is somewhat an outlier – its identification relies not only on overt symptomatology, but also the identification of a particular etiology from traumatic moment to current distress. that is, symptoms diagnosed as ptsd not only index that diagnostic category, but also a previous pathogenic trauma. this indexicality is bolstered by an act of rhematization, the transformation of indexical relationships into iconic links, whereby ptsd symptoms are understood as resembling a pathogenic trauma, thus distinguishing ptsd from disorders whose presentations carry no such resemblance to trauma. keywords: indexicality, rhematization, posttraumatic stress disorder, trauma 1. introduction and history diagnosis is an act of semiosis, transforming abnormal test results and distressing experiences into meaningful signs that point to a specific disease entity. the meaning created by the diagnostician’s interpretation is indexical: as signs, symptoms point to their diagnoses by virtue of being causally linked to them, a conclusion that has been readily drawn by semioticians (eco 1976; ostwald 1964; peirce 1931-1958; sebeok 1994). as a psychiatric disorder uniquely defined by its traumatic cause, posttraumatic stress disorder (ptsd) carries a particularly heavy reliance on indexical and ultimately rhematized links in order to justify itself as a discrete diagnostic entity. in most areas of medicine, conditions are categorized first and foremost on their causes. viral and bacterial bronchitis both cause an unpleasant cough, yet they are distinguished by different pathogens, ultimately leading to different treatment plans. however, psychiatry ultimately rests its categorization schema on groups of observable symptoms. this aspect of modern psychiatry can be traced back to early 20th century psychologist emil kraepelin and is therefore considered the neo-kraepelinian approach (blashfield 2012; compton & guze 1995). under a neo-kraepelinian model, everyone who experiences similar feelings of lethargy, anhedonia, and excessive guilt are prima facie classed together under depressive. this stands in contrast with earlier largely 1 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 psychodynamic approaches that prioritized theoretical underlying causal mechanisms as the basis of categorical generalizations, such that a single cause like weak ego defenses may lead to a range of different presentations while still being theorized as a single entity. while kraepelin imagined that shared causal mechanisms would eventually be identified across shared presentations of symptoms once thusly categorized, for the most part this has yet to happen. the diagnostic and statistical manual of mental disorders (dsm), published by the american psychological association (apa), is for many the ultimate authority of what counts as a mental illness in the united states (and is an influence in many other places of the world).1 it was the third edition of this manual, released in 1980, that truly solidified the neo-kraepelinian approach in what several authors have described as nothing less than a “revolution” (compton & guze 1995; mayes & horwitz 2005; wilson 1993). by emphasizing ostensibly neutral descriptive categories over psychoanalytic etiologies, which could differ between professionals, the psychologists who wrote the dsm-iii hoped to bring a kind of “scientific objectivity” to the field. the dsm-iii asserted itself as an “atheoretical” text whereby “clinicians can agree on the identification of mental disorders on the basis of their clinical manifestations without agreeing on how the disturbances come about” (apa 1980:7). the exception, highlighted in the introduction, was for disorders for which “the etiology or pathophysiological processes are known” (6), such as the organic mental disorders diagnosed after identification of a specific neurological abnormality. the dsm-iii was also noteworthy for introducing a range of new terminologies and diagnostic categories, not simply reorganizing and redefining preexisting ones. one of these was posttraumatic stress disorder (ptsd), a diagnosis whose inclusion largely came from political pressure by war veterans (scott 1990). advocates argued that doctors could only properly treat these veterans if they formally recognized the etiological impact of trauma, and that the “misdiagnosis” of veterans with disorders such as bipolar and schizophrenia was a grave injustice. consequently, at the same time as wide swathes of the dsm were written to avoid suggesting specific causes in the name of being “atheoretical,” we also see the creation of a disorder specifically defined by its theoretical causes. 1 outside of the united states, the world health organization’s (who) international statistical classification of diseases and related health problems (icd) acts as the authoritative text on psychiatric nosology. however, there have been efforts by both who and apa to reconcile the two systems. 2 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 through the lens of indexicality (silverstein 1976) and rhematization (irvine and gal 2000), i suggest that for ptsd to claim legitimacy within a psychiatric framework purporting pure descriptivism, causes must be made just as observable as depressed or anxious affects. this may be done by reading symptoms as unproblematic indexes not just of diagnostic category, the way hallucinations may index schizophrenia, but as rhemes that both point to and phenomenologically resemble a specific trauma. without an etiological cause, a “purely descriptive” rendering of ptsd overlaps considerably with other diagnostic categories: the avoidance of triggering material, mood instability, and even flashbacks may be described as phobic, mood disordered, and psychotic respectively (brewin et al. 2009; mchugh & treisman 2007; young 1995). it is the ability to read a particular causal chain from them that transforms these into symptoms of ptsd. 2. indexes, rhemes, and causes indexicality describes a sign, linguistic or otherwise, that relies on contextual information in order to be interpretable. the current wide use of indexicality can be traced back to charles perice’s tripartite system of types of signs, split by their relation to their signified object: icons, indexes, and symbols. icons are signs whose relation to their signified is one of resemblance or shared qualities, as when a drawn stick figure signifies a human being. symbols are signs whose meanings are purely a matter of convention and which are otherwise wholly arbitrary, such as how the sounds in the word dog mean a particular class of mammal. the relation that defines an index is an existential one: an index “refer[s] to the object that it denotes by virtue of being really affected by that object” (peirce 1931-1958:cp 5.248). this may include physical or spatial contiguity, which directs the interpreter’s attention towards the signified object—such as when a pointed index finger directs attention towards a particular location in space—or, potentially, causal links. an important aspect of this definition is that the interpretation of indexical signs requires an assumption that there is something that exists external to the utterance. smoke indexes fire not because of an arbitrary symbolic relationship, but because the chain of causation from smoke to fire proceeds (in theory) regardless of semiotic interpretation. this connection to a world beyond the sign results in a reliance on contextual information. indexicality has then been taken up as a way to talk about context-dependent meanings, exemplified in deictic expressions such as here, now, i, and you. without sufficient information about the context of the utterance—namely, when it was uttered—the word now holds little 3 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 meaning. however, context is a commonly used word without a clear, uniform definition. common uses highlight its co-constitutive importance to text: context is not the “focal point” or “figure,” but rather the “everything else” that allows for successful interpretation (goodwin & duranti 1992). while this may result in a slightly circular definition—indexicality is that which requires external context, while context is that which is required by indexicality—there is an important insight here: due to the endless possibility of language, everything can be recruited as potential context. think of all the locations that here could possibly refer to! it is the indexical reference itself that transforms inert facts about the world into salient, necessary, contextual information—into socially relevant objects that can be discussed and reasoned about. attention to indexicality helped move the study of meaning away from pure referential semantics, which takes up the study of relatively context-free sentences such as the sky is blue, towards a study of language that necessitates acknowledging its embeddedness in a particular, and a particularly social, context. it is for this reason that indexicality came to be most strongly associated with disciplines that joined language and social reality. silverstein (1976) is credited with bringing the concept of indexicality to linguistic anthropology. he described two actions that an indexical utterance can accomplish: presupposing and performing. because they presuppose some particular context, the indexical deictic terms this and that are uninterpretable and nonsensical if the relevant context is not known; the presupposition fails and meaning breaks down. performative indexicality, also called creative or entailing indexicality, likewise relies on contextual information, but also “seem[s] to be the very medium through which the relevant aspect of the context is made to ‘exist’” (silverstein 1976:34). objects which can be referred to as this or that generally in some way precede their entrance into discourse; however, indexes such as honorifics, which nevertheless require the preexisting context of social convention and a power dynamic among speakers, also enact these social relations in being uttered. when these sorts of indexes “fail,” rather than being referentially baffling, they often change the nature of their medium: honorifics spoken in the wrong contexts may be taken as an insult—a very different kind of social action than the same forms spoken in a different context. these latter indexes fit austin’s (1962) understanding of performativity: they construct their context while simultaneously reflecting it. rather than constituting two separate classes of indexicality, presupposed and creative indexicality exist along a range of possibilities. many indexical utterances can be found to both 4 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 presuppose and perform. silverstein gives the example of first-person pronouns: like other deictics, they presuppose the existence of a person or persons who precede entrance into the discourse—without this context, they are incomputable. however, as first-person pronouns generally mean something like “the speaker of this utterance,” they also create the necessary discourse environment and the notion of a speaking i at all. along the same lines, the word we just as much refers to a preexisting collection of people as it constructs or reinforces a shared group identity. silverstein’s ideas were further refined through the proposal of an indexical order to explain how indexical relations are formed and manipulated in the moment (silverstein 2003). this ordering begins with first-order indexicality: the utterance specific non-referential meanings that come from using one linguistic variant over another. second-order indexicality refers to how this variation is expanded into ideological, metapragmatic meanings as communities attempt to rationalize why linguistic variation exists. a linguistic feature that is recognized by speakers as associated with a certain demographic or location (first-order) can come to signify a certain “type of person” under a particular language ideology (second-order). the word y’all may be statistically more common in the southern united states; however, the social identities attributed to the word, such as “rural,” are second-order indexical meanings that can then be drawn upon when anybody, not just those in the american south, say y’all. this ordering turns on itself indefinitely as speakers draw from second-order indexical meanings, changing the speaker demographics of the linguistic form and generating new metapragmatic understandings: “for any indexical phenomenon at order n, an indexical phenomenon at order n + 1 is always immanent, lurking in the potential of an ethnometapragmatically driven native interpretation” (silverstein 2003:212). silverstein’s class example deals with the highly marked register of wine connoisseurship. this style draws heavily from ideologies regarding what “well-bread,” “expert,” and “high-class” persons sound like. by “wine-talking,” speakers index themselves as wine experts—a type of person that is ideologically linked to above-average intelligence, specialized knowledge, and high socioeconomic class. but what determines what “high-class” persons sound like? it is a second-order understanding of the language varieties found among certain populations, namely the white, rich, and educated. as always, however, n + 1st order indexicality is possible: self-conscious, ironic uses of wine-talk create a new kind of person, namely the kind of person who uses wine-talk ironically. indexical 5 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 ordering then functions as a dialectic, a back-and-forth shifting of language use and language ideology that changes just as soon as speakers develop a metapragmatic understanding of their language variation. for this reason, the terms nth order and n + 1st order are preferred to firstand second-order indexicality—there is no easy way of assessing which (if any) aspects of speaker variation truly came first, absent of any ideological baggage. peirce identified causality as a potential source of indexical relations, specifically calling these causality-based indexical signs reagents. an early example of this was the weathercock, whose position is dictated by the direction of the wind and which therefore indexes this causal force (peirce 1931-1958:cp 2.286). indexicality, by pointing towards pre-existing reality, situates itself in a material world of objects acting upon other objects; this attention to materiality has been described by keane (2003) as a way to “open up signification to causality” (417). however, while peirce seemed to imagine these relations as belonging to the world “irrespective of the interpretant” (cp 2.92), keane noted that there must be ideologies in play that make certain connections recognizable at all. he named these semiotic ideologies: “basic assumptions about what signs are and how they function in the world” (419). these ideologies seem obvious and natural to its holders, and when asked, people can usually give some form of internally-logical explanation for their beliefs. it should be natural, for example, that men should swear more: it’s because of testosterone, innate aggressiveness, socialized competitiveness, and so on. swearing therefore not only indexes masculinity by virtue of convention or mere metapragmatic recognition—i.e., “that’s just the way men speak” without any further value judgements—but seems to on some level resemble or reflect masculinity’s supposed innate aggression and violence. irvine and gal (2000) would describe this as an instance of iconization, later renamed rhematization.2 discussing the rhematization of certain linguistic forms as both indexing and resembling (that is, iconifying) their ideological referents, they described a transformation of the sign relationship between linguistic features (or varieties) and the social images with which they are linked. linguistic features that index social groups or 2 gal’s (2005) recasting of this process as rhematization draws from peirce’s (1940) definition of a rheme as “a sign which, for its interpretant, is a sign of qualitative possibility, that is, is understood as representing such and such a kind of possible object” (103) – in other words, a sign that is interpreted to have qualities (qualia) which are shared with and thus represent some object. 6 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 activities appear to be iconic representations of them, as if a linguistic feature somehow depicted or displayed a social group’s inherent nature or essence (37). that is, the merely indexical relationship between a speech variety and a social identity may be rendered as somehow iconic, such that qualities of language are seen to resemble qualities ascribed to said social group. 3. identifying trauma the diagnosis of ptsd recontextualizes past events into instances of trauma and present cognition or emotions into symptoms of ptsd. it is a potentially paradigm-shifting heuristic through which sufferers and helpers can reinterpret the past and present, imbuing a person’s actions and feelings with a newfound ability to index a traumatic event. that is, there is a sense in which the presently constructed narrative of ptsd grants past events the ability to cause. cause and effect blur further when it is the presence of symptoms that lead clinicians to posit a diagnosis of ptsd even in the absence of clearly-identified traumatic memories (e.g. bass & davis 1988; laibow & laue 1993). young (1996:97-98) goes as far as suggesting that chronic cases of ptsd can be explained just as plausibly if we supposed that time is moving in the opposite direction, that is, from the present (symptoms) back to the past (event). in this scenario, diagnosable depression and anxiety disorders precede the onset of ptsd symptomatology (rather than following or simply co-occurring), and individuals rediscover and rework their memories of past events as a means of accounting for their present distress. this is a troubling scenario, if the nature of indexical links relies on time flowing from cause to event. if indexes are meant to point towards their causes, then this is a form of indexical inversion, a term used by inoue (2004) to describe how the ethno-metapragmatic explanations that define n + 1st order indexical might point to a history—specifically, a linguistic history—that never truly existed. she examined how the idea of women’s language in japan was formed again and again in order to lament its destruction: “the birth of women’s language was also the birth of the corruption of women’s language” (50). it did not necessarily exist even as a statistical correlation until the metapragmatic discussions about the denigrations of women’s language presupposed it into existence. the “corruption” of women’s speech in the late 19th and early 20th was ascribed varyingly to low-class neighborhoods in tokyo, geisha, the mixing of social classes 7 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 in high schools, and contact with westerners. regardless of proposed origin, the ultimate meanings of this corruption were widely agreed upon: sloppiness, laziness, and vulgarity. by suggesting a degree of indexical inversion, i do not mean to suggest that the memories and their effects are invented wholesale; however, their structuring and scaffolding under the framework of trauma relies on the naming of current symptoms as posttraumatic and on the historically and culturally specific science of trauma—in other words, the present-day ideologies that dictate what is defined as trauma. the historically situated nature of these recastings is perhaps clearest in the case of child abuse, which has undergone numerous shifts in meaning over the past century (hacking 1991). events which at the time were not considered abusive or traumatic—were not objects that could be easily indexed by symptoms—become in hindsight instances of child abuse and neglect. an anxiety surrounding the diagnosis of ptsd, then, is ensuring that this powerful indexicality is granted fairly—that there exists a true object being indexed and that the causality supporting the indexical relation is an accurate narrative. causation, after all, evokes blame: the identification of a particular event as causing ptsd opens up the possibility of legal recourse on the basis of psychological injury (day & hall 2016; miller 2015). military veterans who can trace their symptoms to ptsd from a wartime trauma can access disability benefits and resources more easily than if their distress was due to an ostensibly non-traumagenic disorder such as schizophrenia (ray 2014). a seemingly straightforward tactic for determining an adequate causal chain is to ensure that the problematic symptoms only arose after the traumatic event in question. however, even this may prove difficult: ptsd is an appropriate diagnosis if trauma only exacerbates preexisting psychiatric symptoms, and there exists the classification of delayed-onset ptsd when symptoms arise months or years after the traumatic event. additionally, while temporal ordering can suggest causality, it is rarely sufficient as a convincing narrative—indeed, a series of events one after another may be better thought of as “chronicles” than “full-fledged narratives” (carroll 2001:25). ptsd by its nature requires more than a mere chronicle of events; psychiatrists need stronger causal links. young (1995) provided multiple cases of doctors debating the truth of their veteran patients’ etiologies. if there were potential traumatic events prior to the wartime events in vietnam, then there existed a promising alternative narrative wherein the patient’s ptsd was only caused by 8 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 these earlier events; therefore, it was not the obligation of the military hospital to provide treatment. these doctors also struggled with articulating the clear moral evaluation that would point towards a stronger narrative: trauma related to atrocities that the soldier was himself responsible for or personally carried out—so-called perpetrator trauma—complicated any depiction of the veteran as a blameless victim of circumstance and casted suspicion on claims of ptsd. this debate has resurfaced in more recent years centered on cases of ptsd claimed by american drone operators, who are not only viewed as perpetrators but also have to reckon with the difficulties of proving causation across such large distances (mccammon 2017; press 2018; for further discussions on the continuing controversy over perpetration as a source of traumatic stress, see drescher et al. 2011). a common struggle was—and continues to be (breslau et al. 2002)—the identification of as specific a traumatic moment as possible. it was not sufficient to suggest that war in general was traumatic, even when veterans themselves suggested this more diffuse understanding of causality; doctors sought out information about particular battles and even homed in on what event with in a battle constituted “the trauma.” indeed, the idea of a posttraumatic disorder caused by multiple different events over a span of time, as might happen within dysfunctional family dynamics, has been proposed through the alternative construct of “complex posttraumatic stress disorder.” after some debate, complex ptsd was ultimately omitted from the most recent edition of the dsm, with skeptics highlighting how research had yet to clearly illustrate the casual mechanisms from chronic traumatization to the wide range of symptoms described by the proposed disorder, which furthermore overlapped considerably with other diagnoses (resick et al. 2012; şar 2011). in other words, complex ptsd lacked a good, clear causality. 4. rhematizing trauma to help determine what event to highlight and ultimately treat as the pathogenic trauma, doctors and patients turn not only towards the presence or absence of certain symptoms, but also their content. symptoms such as nightmares, flashbacks, and anxiety should all point to the same event, otherwise a diagnosis of ptsd would seem inappropriate (breslau et al. 2002). certain symptoms of posttraumatic stress disorder are seen as not only pointing to the etiological event, but furthermore somehow resembling it. flashbacks are “a dissociative state during which aspects of a traumatic event are reexperienced as though they were occurring at that moment” (apa 9 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 2013:821). the images and sensations of a flashback are similar to—if not indistinguishable from—those same images and sensations experienced during the causal trauma. in much of the popular trauma literature, flashbacks are viewed as a privileged form of remembering (antze 1996; hacking 1995) and perhaps even unique to posttraumatic disorders (brewin et al. 2009). under this ideology, nontraumatic memories are prone to distortion through narrativization, reconstruction, and revision; traumatic memories, on the other hand, and particularly those experienced through disorienting flashbacks, are in some way frozen in time. in van der kolk’s (2013) best-selling book about trauma, the body keeps the score, he writes stories change and are constantly revised and updated … such autobiographical memories are not precise reflections of reality; they are stories we tell to convey our personal take on our experiences (177). in contrast, the imprints of traumatic experiences are organized not as coherent logical narratives but in fragmented sensory and emotional traces: images, sounds, and physical sensations (178). isolated from narrative, such “imprints” can then “be re-experienced without appreciable transformation months, years, or even decades after the actual event occurred” (van der kolk 2002:57, emphasis added). narrative memories can index, in the sense that they point to previous events; traumatic memories iconify, and just as the rhematization of certain linguistic features reifies the “naturalness” of their associations with specific social identities, rhematization makes natural and self-evident the causal links between trauma and symptomology. what allows for this rhematization? hacking (1995), in his analysis of the similarly posttraumatic condition of dissociative identity disorder, coined the term memoro-politics to refer to the disciplining of memory into an object of knowledge. paralleling foucault’s (1978) biopolitics and anatomo-politics, memoro-politics involves the control over the right and wrong was to have memories, structure one’s biography, and tell stories about oneself. these amount to what hacking called “the sciences of the soul,” taking soul to involve “character, reflexive choice, and self-understanding, among much else” (215) and what i may reconfigure as respectively morality, agency, and the reflexive presentation of self, situated within an ideology of what memory is and how it functions. the changing sciences of child abuse have changed what it means to have a childhood: not only can one rename past events as trauma, but the possibility of repressed 10 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 memories creates new narrative arcs and identities (such as the survivor with repressed memories) and the language of flashbacks provides a new form of evidence to support the causal link between traumatic event and symptoms—a particular form of indexical iconicity. it is helpful here to consider other instances of perceiving images or sensations that others do not perceive: hallucinations, particularly those associated with psychosis. before posttraumatic stress disorder was an available diagnosis, many veterans were labelled schizophrenic partially on the basis of hallucinatory experiences, which at times may even include the prototypically psychotic auditory verbal hallucinations (crompton et al. 2017; mccarthy-jones & longden 2015; scott 1990; young 1995). the dsm-5 outlines the need to differentiate between hallucinations and flashbacks, advising that “flashbacks are distinguished from these other perceptual disturbances [hallucinations] by being directly related to the traumatic experience and by occurring in the absence of other psychotic or substance-induced features” (apa 2013:286). this direct relation seems to mean an iconic one, as studies have found individuals diagnosed with schizophrenia whose hallucinatory experiences share only “thematic” or “indirect” relations to trauma (hardy et al. 2005; mccarthy-jones & longden 2015; morrison et al. 2003). without content that can be understood as iconic of a pathogenic trauma, psychotic hallucinations are not understood as flashbacks and will likely not be viewed as caused by trauma at all. 5. conclusion: curing causality all diagnosis is a matter of giving meaning to symptoms reported by or observed in a patient. this interpretation is largely indexical in nature, as symptoms point towards their causes. posttraumatic stress disorder holds a particularly noteworthy tie to indexicality due to its uneasy status under the neo-kraepelinian philosophy of modern psychiatry, which privileges empirical observations over theorized causes. unlike a disorder like schizophrenia, which relies on the identification of present-day symptoms as kinds of hallucinations or delusions, ptsd requires the articulation of a certain type of background—a pathogenic trauma—before symptoms can be read as properly connected to ptsd. therefore, if symptoms can be made to index a pathogenic trauma with the same self-evidence that anxiety can index an anxiety disorder, then posttraumatic disorder remains a coherent category. on the other hand, that fact that ptsd symptoms overlap with those of other disorders has led some scholars to question whether there is anything truly unique about traumatic events and whether arguably nontraumatic events could lead to the same presentation, 11 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 which would potentially undermine the very existence of ptsd and a distinct diagnosis and trauma as a concept (bodkin et al. 2007; brewin et al. 2009; mcnally 2003). in defense of causality, doctors and patients draw from rhematized relationships between symptoms and a patient’s past. finally, ptsd’s strong ties to rhematization can be found in its treatment. one dominant treatment philosophy lies in the proposed difference between traumatic and nontraumatic memory as fragmented iconicity and properly ordered narrative respectively. treatment methodologies such as eye-movement desensitization and repressing (emdr), a form of psychotherapy recommended by the u.s. department of veterans affairs and the american psychiatry association, exist to help the patient integrate traumatic memories into their life story, so that “the memory is transformed into a symbolic verbal account … an autobiographical narrative memory of traumatizing events” that no longer phenomenologically resembles the original traumas (van der hart et al. 2006:319; see also fisher 2014; shapiro 2001; van der kolk 2013). once the processing of remembering is made less iconic, the strict, pathological causal links from past to present slowly unravel. symptoms abate or become more manageable until the patient’s day-today life finally no longer carries rhemes of trauma. references antze, paul. 1996. telling stories, making selves: memory and identity in multiple personality disorder. tense past: cultural essays in trauma and memory, ed. by paul antze & michael lambek, 3–24. new york, ny: routledge. american psychological association (apa). 1980. diagnostic and statistical manual of mental disorders, 3rd ed. washington dc: american psychiatric publishing. apa. 2013. diagnostic and statistical manual of mental disorders, 5th ed. washington dc: american psychiatric publishing. austin, j. l. 1962. how to do things with words. oxford: oxford university press. bass, ellen and laura davis. 1988. the courage to heal: a guide for women survivors of child sexual abuse. new york, ny: harpercollins. blashfield, roger k. 2012. the classification of psychopathology: neo-kraepelinian and quantitative approaches. new york, ny: springer science & business media. bodkin, j. alexander; harrison g. pope; michael j. detke; james i. hudson. 2007. is ptsd caused by traumatic stress? journal of anxiety disorders 21.176–182. 12 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 breslau, naomi; g. a. chase; and j. c. anthony. 2002. the uniqueness of the dsm definition of post-traumatic stress disorder: implications for research. psychological medicine 32.573– 576. brewin, chris r.; ruth a. lanius; andrei novac; ulrich schnyder; and sandro galea. 2009. reformulating ptsd for dsm-v: life after criterion a. journal of traumatic stress 22.366– 373. carroll, noël. 2001. on the narrative connection. new perspectives on narrative perspectives, ed. by willie van peer and seymour benjamin chatman, 21–42. new york, ny: state university of new york press. compton, wilson m. and samual b. guze. 1995. the neo-kraepelinian revolution in psychiatric diagnosis. european archives of psychiatry and clinical neuroscience 245.196–201. crompton, laura; yael lahav; and zahava solomon. 2017. auditory hallucinations and ptsd in ex-pows. journal of trauma & dissociation 18.663–678. day, terri r. and rycan c. w. hall. 2016. ptsd and tort law. comprehensive guide to posttraumatic stress disorders, ed. by colin r. martin; victor r. preedy; and vinood b. patel, 231–244. new york, ny: springer. drescher, kent d.; david w. foy; caroline kelly; anaa leshner; kerrie schutz; and brett litz. 2011. an exploration of the viability and usefulness of the construct of moral injury in war veterans. traumatology 17:8–13. eco, umberto. 1976. a theory of semiotics. bloomington, in: indiana university press. fisher, janina. 2014. the treatment of structural dissociation in chronically traumatized patients. trauma treatment in practice: complex trauma and dissociation, ed. by trine anstorp and kirsten benum. traumebehandling. komplekse traumelidelser og dissosiasjon [trauma treatment in practice: complex trauma and dissociation.] oslo: universitetsforlaget. foucault, m. 1978. the history of sexuality, vols. 1, 2. new york, ny: pantheon. gal, susan. 2005. language ideologies compared. journal of linguistic anthropology 15.23–37. goodwin, charles and alessandro duranti. 1992. rethinking context: an introduction. rethinking context: language as an interactive phenomenon, ed. by charles goodwin and alessandro duranti, 1–42. cambridge ma: cambridge university press. hacking, ian. 1991. the making and molding of child abuse. critical inquiry 17.253–288. 13 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 hacking, ian. 1995. rewriting the soul: multiple personality and the sciences of memory. princeton, nj: princeton university press. hardy, amy; david fowler; daniel freeman; ben smith; craig steel; jane evans; philippa garety; elizabeth kuipers; paul bebbington; and graham dunn. 2005. trauma and hallucinatory experience in psychosis. journal of nervous and mental disease 193.501–507. inoue, miyako. 2004. what does language remember? indexical inversion and the naturalized history of japanese women. journal of linguistic anthropology 14.38–56. irvine, judith t. and susan gal. 2000. language ideology and linguistic differentiation. regimes of language: ideologies, polities, and identities, ed. by paul v. kroskrity, 35–84. santa fe, ca: school of american research press. keane, webb. 2003. semiotics and the social analysis of material things. language & communication 23.409–425. laibow, rima e. and c. shaffia laue. 1993. posttraumatic stress disorder in experienced anomalous trauma. international handbook of traumatic stress syndromes, ed. by john p. wilson and beverly raphael, 93–103. boston, ma: springer. mayes, rick and allan v. horwitz. 2005. dsm-iii and the revolution in the classification of mental illness. journal of the history of the behavioral sciences 41.249–267. mccammon, sarah. 2017. the warfare may be remote but the trauma is real. national public radio, 24 april 2017. online: https://www.npr.org/2017/04/24/525413427/ mccarthy-jones, simon and eleanor longden. 2015. auditory verbal hallucinations in schizophrenia and post-traumatic stress disorder: common phenomenology, common cause, common interventions? frontiers in psychology 6. mchugh, paul r. and glenn treisman. 2007. ptsd: a problematic diagnostic category. jouranl of anxiety disorders 21.11–222. mcnally, richard j. 2003. progress and controversy in the study of posttraumatic stress disorder. annual review of psychology 54.229–252. miller, laurence. 2015. ptsd and forensic psychology: applications to civil and criminal law. new york, ny: springer. morrison, anthony p.; lucy frame; and warren larkin. 2003. relationships between trauma and psychosis: a review and integration. british journal of clinical psychology 42.331–353. 14 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 ostwald., peter f. 1964. how the patient communicates about disease with the doctor. approaches to semiotics, ed. by thomas a. sebeok; alfred s. hayes; and mary catherine bateson, 11–34. the hauge: mouton & co. peirce, charles s. 1931–1958. collected papers of charles s. peirce, vol. 2, ed. by charles hartshorne and paul weiss. cambridge, ma: harvard university press. peirce, charles s. 1940[1897, 1903, 1910]. logic as semiotic: the theory of signs. philosophical writings of peirce, ed. by justus buchler, 98–119. new york, ny: dover. press, eyal. 2018. the wounds of the drone warrior. new york times, 13 june 2018. online: https://www.nytimes.com/2018/06/13/magazine/veterans-ptsd-drone-warrior-wounds.html ray, christopher l. 2014. feigning screeners in va ptsd compensation and pension examinations. psychological injury and law 7.370–387. resick, patricia a.; michelle bovin; amber calloway; alexandra dick; matthew king; karen mitchell; michael suvak; stephanie wells; shannon wiltsey stirman; and erika wolf. 2012. a critical evaluation of the complex ptsd literature: implications for dsm-5. journal of traumatic stress 25.241–251. şar, vedat. 2011. developmental trauma, complex ptsd, and the current proposal of dsm-5. european journal of psychotraumatology 2.1–9. scott, wilbur j. 1990. ptsd in dsm-iii: a case in the politics of diagnosis and disease. social problems 37.294–310. sebeok, thomas a. 1994. signs: an introduction to semiotics. london, uk: university of london press. shapiro, francine. 2001. eye movement desensitization and reprocessing (emdr), 2nd ed. new york, ny: the guilford press. silverstein, michael. 1976. shifters, linguistic categories, and cultural description. meaning in anthropology, ed. by keith h. basso and henry a. selby, 11–55. albuquerque, nm: university of new mexico press. silverstein, silverstein. 2003. indexical order and the dialectics of sociolinguistic life. language & communication 23.193–229. van der hart, onno; ellert r. s. nijenhuis; and kathy steele. 2006. the haunted self: structural dissociation and the treatment of chronic traumatization. new york, ny: w. w. norton & company. 15 parish: rhematization as etiology in the diagnosis of ptsd published by cu scholar, 2019 van der kolk, bessel. 2013. the body keeps the score: brain, mind, and body in the healing of trauma. new york, ny: penguin books. wilson, mitchell. 1993. dsm-iii and the transformation of american psychiatry: a history. the american journal of psychiatry 150.399–410. young, allan. 1995. the harmony of illusions: inventing post-traumatic stress disorder. princeton, nj: princeton university press. 16 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/6 doi: http://dx.doi.org/10.33011/cril.24.1.6 colorado research in linguistics 6-2019 rhematization as etiology in the diagnosis of posttraumatic stress disorder ayden parish recommended citation rhematization as etiology in the diagnosis of posttraumatic stress disorder cover page footnote microsoft word parish-cril2019-final.docx microsoft word raclaw-cril2021_proof-final.docx 1 “i saw you like this now i wanna know”: noticing recipiency and responding to likes on twitter joshua raclaw, lauren durante, olivia marquardt west chester university in this paper, we focus on how interactants accomplish different forms of participation in the “one-tomany” context of social media interactions, where single users frequently have a wide audience of potential recipients to their posts. how do social media users ascertain who might be a relevant recipient to these posts, and how do other users who interact with these posts position themselves within a relevant participation framework? we explore these questions by examining how participants on twitter orient to the act of “liking” a post as a resource for moving into the participation framework of the talk, and we show how this orientation allows likes to serve as possible pathways for launching new actions and activities. we examine these practices using the framework of conversation analysis (ca), showing how participants use public noticings of another user’s likes as a preface to, and justification for, a subsequent invitation sequence or complaint sequence. we additionally show how the specific media affordances of twitter, which render likes publicly visible to others, facilitates the noticing of likes as a device for inciting new courses of action. keywords: conversation analysis, social media, response practices, twitter during face-to-face interaction, participants routinely display their recipiency to an ongoing turn at talk through a range of interactional practices that include minimal response tokens or backchannels (goffman, 1981) and embodied resources such as eye gaze and head nods (goodwin 1980; stivers, 2008). such practices treat recipiency as an accomplishment that requires work beyond simply being physically co-present with one’s interlocutors, though co-presence is also, in and of itself, a valuable resource for participation (goodwin & goodwin, 2004). in the context of face-to-face interaction, being visibly co-present enables some of the most basic turn-taking mechanisms of interaction (sacks, schegloff, & jefferson, 1974) by allowing current speakers to address ongoing actions to particular participants, and it additionally allows participants to position themselves in meaningful ways within the embodied participation framework (goodwin, 2000) of the talk. the interactional affordances of visible co-presence are not only made relevant during faceto-face interactions but also those conducted over videotelephony platforms like zoom or facetime, where participants can keep their cameras on to maintain a visible but remote cocolorado research in linguistics, volume 25 (2021) 2 presence. however, interaction occurring in other technologically mediated environments that lack the affordances of visible co-presence, such as telephone calls or text-based modes of digital communication, may motivate shifts in how interactants project, invite, and enact participation in the talk. in this paper, we focus on how interactants accomplish different forms of participation in the “one-to-many” context of social media interactions, where single users frequently have a wide audience of potential recipients to their posts. how do social media users ascertain who might be a relevant recipient to these posts, and how do other users who interact with these posts position themselves within a relevant participation framework? we explore these questions by examining how participants on twitter orient to the act of “liking” a post as a resource for moving into the participation framework of the talk, and we show how this orientation allows likes to serve as possible pathways for launching new actions and activities. we examine these practices using the framework of conversation analysis (ca), extending prior work that has used ca to examine various forms of text-based digital communication (e.g., meredith, 2019) but has only rarely focused on social media as a site for conversation analytic inquiry. one notable exception to this trend is housley et al.’s (2017) exploration of twitter data using ethnomethodological conversation analysis, which focuses on both sequential organization and membership categorization devices. in a somewhat similar vein, giles (2021) offers an in-depth discussion of sequential organization on twitter that interrogates platform-specific issues of context and media affordances. giles additionally illustrates the relevance of doubly articulated units of talk (conceived with both a local and wider audience in mind; bou-franch et al., 2012) to a micro-analytic account of talk on twitter, particularly interactions between public celebrities and the fans who follow them. while both of these papers focus on the written modality of social media interactions, other streams of discourse analytic research have investigated the interactional work that participants accomplish by liking posts on social media. west and trester (2013) offer one such discussion of likes on facebook, which they describe as a response practice “signaling acknowledgment and approval” (p. 138) of a post’s content. the authors focus on the types of facework conducted by responding to posts with written comments, describing such comments as a form of “meaningful engagement” that contrast with the more limited possibilities for doing facework engendered by simply liking a post, which instead offers only a “minimal effort response” (p. 145). their analysis positions likes as working toward pro-sociality by offering quick, positive feedback, yet nonetheless carries the “risk” of being a missed opportunity for the “i saw you like this now i wanna know” 3 types of facework made available through leaving comments on a user’s post. in a briefer discussion of likes on facebook, west (2013) describes them specifically as a backchannel device (goffman, 1981) and again contrasts them with written comments left on a social media post, which are instead described as an “active” form of response practice. though interaction and interpersonal engagement on facebook differs from twitter in significant ways, the two platforms overlap considerably in how both offer two primary ways of responding to a user’s post: by liking it or producing written comments. (while facebook has expanded its platform in recent years to offer users a range of emotive “reactions” in addition to likes, this feature was not yet implemented during the prior research cited here). west and trester (2013) and west (2013) thus both offer relevant insight into the ways that likes may be used and understood across twitter as well as facebook. more recent work by proctor and raclaw (2018) has focused on the social meaning of likes on twitter, examining how users produce metacommentary (i.e., talk about talk) about the presence or absence of likes on their posts. the authors find that likes are treated as a noticeable form of response, with participants producing explicit noticings that either celebrate or lament the likes a particular post has received. it is this particular understanding of likes as a noticeable form of response that drives much of the present analysis, which applies a conversation analytic lens to examine participants exploit this noticeability to incite further forms of participation in interactions on twitter. in particular, we show how participants use public noticings of another user’s likes as a preface to a new course of action, such as an invitation or complaint. in this way, likes are treated as providing for another user’s availability as a relevant recipient to a subsequent unit of talk. for example, in excerpt 1, sam posts a single tweet at lines 1-4. the tweet is a humorous announcement formulated using a popular compound tcu (lerner, 1991) meme format that contrasts sam’s inner monologue about what they should do for the evening (stay at home because they have work the following morning) and what they actually spent their evening doing (going out dancing). colorado research in linguistics, volume 25 (2021) 4 (1) 01 sam: me to myself: i am not going out tonight 02 i have work in the morning 03 me the same night: 04 ((animated gif of child dancing in a nightclub)) 05 sam: linda i saw you like this, come hop state lines 06 with the crew 07 lin: i’m in br for landons state tournament :/ 08 sam: i’ll take shot for you and him xoxo this initial tweet receives several likes, including one from linda, who is addressed as a recipient in sam’s subsequent talk at lines 5-6. here sam formulates a noticing of the fact that linda had liked the initial tweet before producing an invitation for linda to “hop state lines with the crew,” possibly to participate in the same activity described in the initial tweet (going out dancing). linda rejects the invitation at line 7 by offering an account for why they are in fact unavailable for a visit, and sam closes the sequence at line 8 by accepting the rejection (“i’ll take [a] shot for you and him”). in this excerpt, we see that linda’s like of the initial tweet is treated as moving them into the participation framework of the talk by positioning them as a relevant recipient to sam’s invitation. this particular understanding of linda’s like is made salient during sam’s explicit noticing of this like at line 5, which is formulated as both a preface to, and justification for, sam’s subsequent invitation. we note that this understanding of likes as a springboard for a new course of action is in part enabled through one of the media affordances (giles, 2018) of twitter, namely in how the platform automatically notifies the author of a post or comment when another user has liked it. even outside of these notifications, likes are publicly visible to other users who come across the original post or comment, and twitter’s timeline algorithm may even show users which tweets have been liked by other users they follow. this relatively high, public visibility of likes facilitates subsequent turns at talk in which these likes can become explicitly noticed. such noticings are routine occurrences in the data we examine, and they typically serve as both a preface to, and justification for, some new course of action that unfolds in the talk that follows. “i saw you like this now i wanna know” 5 a related case occurs in excerpt 2 as arc posts a single tweet at lines 1-7. this tweet is composed of multiple units of talk: an initial instance of troubles talk (jefferson, 1988) about a problem in the game of dungeons and dragons that arc runs, followed by a solicitation of advice formulated through two questions. in terms of recipiency, the tweet is directed to a limited but still potentially vast set of recipients, namely individuals who also run games of dungeons and dragons (serving as a dm or gm, respectively short for “dungeon master” and “game master”). (2) 01 arc: 🤔 ok dm/gm friends. an open campaign recently 02 took such a hard left i’ve found myself searching 03 for an idea for a decent arc... and coming up 04 with nothing...it happens... 05 has it ever happened to you? and how did you work 06 through it? 07 ((animated gif of actor nathan fillion)) 08 lor: ask the players what their theories are and adlib 09 off off that, or just do a fun, completely 10 unrelated side arc and see where it leads (man a 11 one-shot is written to fit into any setting :d) 12 arc: that’s great advice! unfortunately it’s that side 13 arc i’m searching for lol. so far the theories 14 haven’t solidified. and sadly, in this case, i 15 would be absolutely amazed if a one-shot actually 16 fit the setting/situation...i’ve really stuck my 17 foot in it 😆 18 arc: i saw you like this shit @shad. you up for a call? 19 in fact, who’s up for a discord voice chat? @chao, 20 @tx, @dust? anyone else? at lines 8-11 lorai responds with advice, and at lines 12-17 arc initially accepts and praises the advice but ultimately rejects it as irrelevant to the trouble at hand. subsequently, at lines 18-19 arc colorado research in linguistics, volume 25 (2021) 6 produces a noticing of shad’s like of the initial post from lines 1-7. just as in the prior excerpt, this noticing is formulated as a preface to, and an account for, an invitation: at lines 18 arc “tags” shad by mentioning their username (which sends a notification to shad alerting them to this tweet) and invites them to talk about arc’s trouble at hand (an invitation that is broadened out to other users at lines 19-20). as with the prior excerpt, shad’s liking of the initial tweet is treated as positioning them as a relevant recipient to a new course of action—an invitation. while likes may thus be used to “indicate having noticed and appreciated a friend’s post” (west & trester 2013:145), the data from our larger collection illustrate how likes on social media may also position a participant as being interested in the talk such that their further participation is made relevant. in the prior two excerpts, the talk is organized such that the original author of a post notices another user’s like and thus initiates the subsequent invitation sequence. in other cases, a third party goes on to produce this noticing as well as the new course of action that the original like has engendered. for example, in excerpt 3 the official twitter account for the multiplayer video game dead by daylight formulates an announcement advertising an unlockable download for players of the game (lines 1-5). the original post does not specify any one recipient, though it receives a response from a user called leila who notices that the original post was liked by the official twitter account for trixie mattel, a celebrity drag queen and television personality. (3) 01 dbd: zarina's bringing in the year of the ox in style. 02 if you want to be like zarina... enter code 03 "zarinox" in the in-game store by february 25th to 04 unlock this limited time lunar new year cosmetic. 05 ((image of the game character zarina)) 06 lei: excuse me @trixiemattel 07 i saw you liked this does it mean you play will 08 you party with me? 🥺 at lines 6-7 leila first tags trixie mattel by mentioning her username, then formulates an explicit noticing of mattel’s like of the original tweet that prefaces leila’s invitation for mattel to join them in a multiplayer game of dead by daylight by forming an in-game “party” (the invitation is “i saw you like this now i wanna know” 7 additionally accompanied by a “pleading face” emoji). while leila’s invitation receives no uptake from mattel, it offers an example of how liking a tweet can be understood as positioning a participant as a relevant recipient to a related course of action (here again, an invitation), even when this noticing is accomplished by a third party rather than the author of the original post. similarly, in excerpt 4, a popular twitter account, rate my takeaway, posts a video of food service workers at the restaurant chip inn preparing a large meat box with curry sauce (lines 1-2). as with the prior excerpt, this initial post does not specify any one recipient, though it receives a response from ben as they notice that the original post was liked by a mutually known party, soph (line 3). (4) 01 rmt: 15" chip inn meat box with curry sauce 02 ((video of service workers preparing food)) 03 ben: @soph i saw you liked this, its 15 mins away 04 from me and its fire 05 sop: omw to yours now 06 ben: its at a place called huthwaite ben’s noticing of soph’s like at line 3 serves as a preface to two subsequent units of talk: an announcement that the restaurant featured in the video is only 15 minutes away from where ben lives, and a positive assessment of either the restaurant or the specific meal advertised in the original post (“it’s fire”). while neither unit of talk formulates an explicit, on-record invitation for soph to visit the restaurant, it is nonetheless heard that way as soph responds at line 5 by announcing that they are “on [their] way” to visit ben, ostensibly so that the two of them might visit the restaurant together. while ben’s subsequent turn at talk (line 6) disaligns with this particular interactional project—that is, it offers soph specific directions to get to the restaurant on their own rather than solidifying plans for the two of them to visit together—this case nonetheless illustrates how soph’s like has positioned them as potentially interested in and available for further participation regarding the content of the original post. ben’s noticing of soph’s like thus becomes a preface to, and an account for, this expanded participation, which soph treats as an invitation. colorado research in linguistics, volume 25 (2021) 8 the previous excerpts each illustrate how likes may be treated as noticeable forms of response that provide for the relevance of the respondent’s further participation in the talk. in each of these cases this call to participation is treated as an invitation. and yet because of their sequential organization, none of these noticings are quite analogous to the types of pre-invitations (schegloff, 2007) that speakers routinely use during talk-in-interaction to first ascertain the relevance of an invitation sequence. for example, in the landline telephone interaction below, nelson initiates a pre-invitation at line 4 as he checks to see whether clara is available for the subsequent invitation that follows at line 6, while clara signals this availability through the “go ahead” response she provides at line 5. (5) 04 nel: whatcha doin’. 05 cla: not much. 06 nel: y’wanna drink? 07 cla: yeah. 08 nel: okay. here, the pre-invitation checks the recipient’s availability for a specific course of action—the invitation. by contrast, in the twitter data examined above, a user’s like does somewhat different work; rather than simply providing for the specific action-type relevance of a forthcoming invitation, these likes provides for the respondent’s more general relevance as a recipient to a subsequent course of action. by explicitly noticing these likes, and organizing such noticings as prefaces to this next course of action, participants display an understanding of likes as signaling both the participant’s interest in the talk as well as their potential availability as a relevant participant within it. though our focus thus far has been on the way that likes can engender a subsequent invitation, our collection also shows how other courses of action may also accompany the public noticing of other participants’ likes. for example, excerpt 6 begins as karti formulates a hyperbolic complaint about mint chocolate chip flavoring and the people who like it (lines 1-2). another participant, mari, follows this at line 3 with a turn composed of three distinct units of talk directed at a third “i saw you like this now i wanna know” 9 party, elli, who has liked this initial tweet: an initial response cry (“what the hell”) followed by a negative assessment (“you tweakin”) and a noticing of elli’s like (“i saw you like this”). (6) 01 kar: if u like mint chocolate chip anything seek 02 help ur going 2 hell 03 mar: @elli wth you tweakin i saw you like this 04 ell: i liked it because mint chocolate chip is my 05 favorite ice cream 😭 06 mar: ohhh i thought you were agreeing 😭😭 i was 07 gonna say you missing out in contrast to the prior cases we have analyzed thus far, mari’s noticing of elli’s like is not organized as a preface to the complaint they launch at elli at line 3, but rather serves as the final unit of talk within her turn. despite this difference in turn construction, mari’s noticing of elli’s like is nonetheless positioned as justification for mari’s complaint and, more precisely, elli’s like itself is positioned as the complainable. at lines 4-5 elli responds by accounting for her like, noting that they liked the original tweet not because they agreed with the stance that it put forward but rather because they do, in fact, like mint chocolate chip (formulated through the extreme case formulation, “mint chocolate chip is my favorite ice cream”). mari responds at lines 6-7 with an initial change of state token that offers an acceptance of this account and a justification for their original complaint from line 3. a similar case occurs in excerpt 7. at lines 1-3, u.s. republican leader kevin mccarthy posts some points of disagreement with the covid financial relief plan that was then being put forward by democratic leadership. at lines 5-10 another user, jess, produces a single tweet responding to mccarthy and disagreeing with his argument that funding for the arts should not be a part of this relief plan. at lines 11-16, jess then produces a subsequent tweet that initially tags their local political representative, senator john cornyn, who has liked mccarthy’s tweet; jess then produces an initial noticing of cornyn’s like. colorado research in linguistics, volume 25 (2021) 10 (6) 01 km: dear democrats: stop calling it a “covid 02 relief” plan. a better name would be “the 03 pelosi payoff.” 04 ((graph comparing covid and non-covid funding)) 05 jes: arts funding is not non-covid. arts and culture 06 are a key & significant part of our economy and 07 job market. & covid has shut it down almost 08 completely. i am an arts marketer & currently on 09 unemployment because i lost my job. because of 10 covid. learn @gopleader. listen. for once. 11 jes: also, @johncornyn i saw you liked this & i’m 12 absolutely disgusted that you “represent” me. i 13 miss the arts. i miss working. i miss my industry. 14 i hate seeing so many of my colleagues and 15 friends who are artists suffering. because of our 16 countries incompetence. this noticing of cornyn’s like is formulated as a preface to jess’s subsequent complaint against cornyn (“i’m absolutely disgusted that you ‘represent’ me”), which is followed by further disagreements with mccarthy’s stance that offer accounts for the complaint against cornyn. much as with the prior excerpt, jess’s noticing of cornyn’s like serves as justification for the complaint that follows, with the like itself serving as the complainable. as seen in the transcript above, cornyn does not respond to this complaint. in both this and the prior excerpt, likes may be understood as not just approving of a stance put forward in the liked tweet (cf. west & trester, 2013) but also espousing this stance. public noticings of these likes are thus positioned as justifying the complaints that call these parties to account for these likes and, by extension, the stances they index. the likes seen in excerpts 6 and 7 thus differ from those seen in excerpt 1-4, with the former being treated as affiliating with the stance put forth in the tweet the participant has liked, and the latter being treated as signaling that the participant is sufficiently interested in the topic of the talk that an invitation is made relevant. “i saw you like this now i wanna know” 11 however, each of these cases illustrate how participants on twitter treat likes as a noticeable form of response that may be used to further bring these respondents into the participation framework of the talk. while likes may in fact be a more “passive” form of response compared to the types of written comments that also abound on social media (west & trester, 2013), likes nonetheless engender “active” forms of participation as other participants treat them as justification for pursuing further courses of action such as invitations or complaints. each of these excerpts also illustrate the way that participants are held accountable for their likes; in this sense, likes are not simply neutral ways of acknowledging a post or comment, but also display various stances toward the content that being liked, with such stances forming the basis for the invitation and complaint sequences that we see unfold in the excerpts above. we note that it is the specific media affordances of twitter, that render likes so publicly visible to others, that facilitates the noticing of likes as a device for inciting these new courses of action. references bou-franch, p., lorenzo-dus, n.,& garcès-conejos blitvich, p. (2012). social interaction in youtube text-based polylogues: a study of coherence. journal of computer mediated communication, 17, 501–521. giles, d.c. (2018). twenty-first century celebrity: fame in digital culture. emerald. goffman, e. (1981). forms of talk. university of pennsylvania press. goodwin, c. (1980). restarts, pauses, and the achievement of a state of mutual gaze at turn-beginning, sociological inquiry, 50(3-4), 272–302. goodwin, c. (2000), action and embodiment within situated human interaction. journal of pragmatics, 32(10), 1489–522. goodwin, c. & goodwin, m. h. (2004). participation. in a. duranti (ed.) a companion to linguistic anthropology (pp. 222–244). blackwell. housley, w., webb, h., edwards, a., procter, r., & jirotka, m. (2017). digitizing sacks? approaching social media as data. qualitative research, 17(6), 627–644. meredith, j. (2019). conversation analysis and online interaction. research on language and social interaction, 52(3), 241–256. schegloff, e. a. (2007). sequence organization in interaction: a primer in conversation analysis, volume 1. cambridge university press. stivers, t. (2008). stance, alignment and affiliation during storytelling: when nodding is a token colorado research in linguistics, volume 25 (2021) 12 of affiliation. research on language and social interaction, 41(1), 31–57. west, l. e. (2013). facebook sharing: a sociolinguistic analysis of computer-mediated storytelling. discourse, context & media, 2, 1–13. west, l. & trester, a. m. (2013). facework on facebook: conversations on social media. in d. tannen & a. m. trester (eds.) discourse 2.0. language and new media (pp. 133–154). georgetown university press. the evolution of evolutionary linguistics colorado research in linguistics. june 2007. vol. 20. boulder: university of colorado. © 2007 by jeff roesler stebbins. the evolution of evolutionary linguistics jeff roesler stebbins university of colorado for more than a century after darwin’s origin of species, linguists said little about the origins of human speech. in the past 30 years, however, some linguists and evolutionary biologists have proposed descriptions of the roles of gestures, the vocal apparatus, cognition, syntax, and social interaction in the emergence of language. this paper summarizes some of their claims, especially those that assume the certainty of neo-darwinian evolution. neo-darwinism, though, has various critics disputing its claims to be settled fact. after brief consideration of some of those criticisms, the paper will encourage linguists to exercise more caution in their dependence upon neo-darwinian theory. finally, several other fields of science will be mentioned as possible candidates for offering linguists an increasing understanding of the emergence of speech. 1. introduction many linguists are aware of the 1866 linguistic society of paris’ ban on discussion of the evolution of language shortly after darwin’s 1859 origin of species (e.g. newmeyer 2003:59). this formal ban spread informally elsewhere, and until recently, silence ruled. it appears that the topic re-emerged a century later after bickerton’s discussion of “proto-language” in his 1981 roots of language, but it could also be that interest in language evolution increased after comments in john lyons’ widely-used two-volume semantics, wherein he wrote, the attitude of most linguists to evolutionary theories of the origin of language tends to be one of agnosticism. psychologists, biologists, ethologists and others might say, if they so wish, that language must have evolved from some non-linguistic signaling-system; the fact remains, linguists might reply, that there is no actual evidence from language to support this belief (1977:85-6). lyons here echoes questions he had raised seven years before in an earlier title (1970:229). regardless of the source of their inception, discussions of the evolution of language have recently proliferated–so much so that there is now a bi-annual international conference on the evolution of language. its sixth meeting was in rome in april of 2006. three decades is not long for any new discipline or sub-discipline, and the field still seems to be in its formative stages. there are ‘evolutionary biologists,’ but ‘evolutionary linguists’ remain hard to find. still, we might reasonably speak of ‘evolutionary biolinguistics,’ for cambridge has published a text entitled biolinguistics, with a chapter on the evolution of language. and tecumseh fitch, 1 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 2 for example, is an evolutionary biologist studying the physiology of animal communication. he often cooperates with marc hauser, whose research is in zoology and physical anthropology. richard dawkins, an evolutionary biologist, and robert pennock, a philosopher of science, have also considered the evolution of language in popular books, but neither are likely to call themselves linguists. it is not clear–nor perhaps need it be–whether the new topic should be a branch of linguistics, of biology, a marriage of the two, or a part of physical anthropology. hauser, chomsky and fitch, in a much-discussed article in science, sought to “promote a stronger connection between biology and linguistics,” and to “clarify the biolinguistic perspective on language and its evolution” (2002:1570). ever since darwin, of course, evolution has been a potent term in the life sciences. other disciplines (such as economics and political science) occasionally appropriate it as a metaphor for developments observed in their fields. those who study the origin and development of language, however, are not merely appropriating evolution as a metaphor; rather, they are applying evolutionary biology to human speech as the foundational approach in which to conduct their research. this paper will therefore summarize what prominent linguists and biologists are writing about the evolution of language before considering the implications of linguists’ dependence upon evolutionary biology. 2. evolutionary linguistics every discipline depends heavily upon clear definitions of its terminology, and evolutionary linguistics may need to do so more than most. writers in this new field argue for certain precursors to (or essential building blocks of) language. assuming that humans are what we are, and have language as we have it, because of long processes of natural selection acting upon random variations, then what do we have, and how and in what order did we acquire it? 2.1. essential building blocks some precursors of speech are obvious, even to lay persons: the abilities to speak and to hear, agreed upon lists of words, and so on. but within linguistics, psycholinguists, phonologists, syntacticians, semanticists and others each emphasize their own respective foci of study, whether they are conceptual frameworks, vocal physiology, systems of reference, or word order. nobody seems able to agree upon the sequence in which these several elements of speech must have evolved, or even if it would have been possible for any of the phenomena to emerge without the simultaneous emergence of all of them. such is the interrelatedness of the ingredients of language that linguists have difficulty imagining any existing independent of most others. the index to jackendoff’s foundations of language, for example, lists 18 interface relationships (in which one element of language interfaces with another): intonation with syntax, phonology with conceptual structures, syntax with semantics and pragmatics, and gestures with morphophonology are just four examples (2002:469). some have 2 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 3 even wondered whether language might actually be an irreducibly complex system, which would necessarily preclude it evolving piecemeal, a bit at a time. lieberman says that “brain mechanisms adapted for adaptive motor control were the starting point for the evolution of human language” (2003:255). after that, lieberman’s sequence is unclear. knight, studdert-kennedy and hurford believe that “the emergence of syntax was the final step” (2000:4). jackendoff agrees (2002:260-1). in the absence of a clear evolutionary sequence, linguists and other scientists have divided their labor among the parts of what they think may have happened, leaving other parts of the sequence to experts in other fields. they may not know what happened when; still, what follows will briefly describe their hypotheses about several ingredients of the evolution of language in this order: primate gestures, the vocal apparatus, cognition and logic, syntax, and the social elements of the development of speech. 2.2. a need for tentative hypotheses in their dependence upon the presumed certainties of evolutionary biology, some in evolutionary linguistics make strong claims with words such as know, certainly, and obviously; others in the field are more circumspect. macneilage and davis, for example, begin their discussion of evolving speech complexity asserting, “it is common sense that speech must have been simpler in earlier times than it is now” (2000:148). but linguists have so far sought in vain for evidence to support that claim, which is disputed by others writing on the topic (e.g. pinker 2003:22). no trace of anything like a ‘primitive’ language has ever been found. macneilage and davis use ‘must have’ four times in five lines. such confidence might be warranted were there certainty in the evolutionary biology upon which they depend, but (as will become clear below) this is problematic. fitch, on the other hand, begins more modestly by saying that “discussions of the evolution of language often involve more speculation than data” (2000:258). in contrast to macneilage and davis, fitch uses might have or could have four times in a dozen lines (2000:263). the following, then, is a partial, tentative outline of what some linguists say might have happened: 2.2.1. primate gestures evolutionary linguistics presupposes the theory that non-human primates and humans are descended from some common, prehistoric ancestral primates. linguists therefore study the behavior of other primates, none of whom share the vocal apparatus or vocal acuity of humans. other primates, however, do use manual and facial gestures communicatively, and humans have trained some to use a few hundred words of sign language. linguists therefore study primate gestures to discover what they may have in common with our non-vocal gestures, which are assumed to have preceded human speech (e.g. hewes 1973:5-24, corballis 2003:201-18). modern human ‘body language’ is also thought by some 3 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 4 to include vestiges or ‘fossils’ of gestural communication used by humans’ and apes’ common ancestors long ago. years of research has shown that, unless they receive extensive training (and most, even if they do), primates do not imitate others’ gestures or vocalizations, do not point, are apparently incapable of directing or sharing attention (which requires a theory of mind–see below), and do not use gestures referentially to represent anything not present (tomasello 2003:100-1, corballis 2003:203). and even when primates (orangutans, chimpanzees, bonobos, or gorillas) receive years of training, they remain unable to understand or use tenses, questions, commands, recursion, or even negation (corballis 2003:204). giacomo rizzolatti and michael arbib believe they may have discovered how we evolved our capacity for imitative gestures and vocalizations. a part of primate brains appears to cause them to grasp (close their hands) when they see another grasping something. because this part shares (or is close to) the part of the brain responsible for vocalizations, rizzolatti and arbib say that this area of the brain may have contained evolutionary ‘bridges’ from mere manual movements to imitative gestures, imitative vocalizations, and then presumably, communicative vocalizations (rizzolatti and arbib 1998, arbib 2003). while various birds (esp. parrots and mynahs) and some aquatic mammals have demonstrated amazing abilities of vocal learning and imitation, primates have proven especially disappointing in this regard (fitch 2000:261); they appear unable to voluntarily control vocal musculature, or even to restrain emotional vocalizations when it would be safer to do so (lieberman 2003:258). fitch, too, points out that other primates lack our “freedom from stimulus-driven control of vocalization” (2000:265). so far, comparisons of human language and primate gestures (whether vocal or non-vocal) appear to teach us more about how we differ than about what we may have in common. experiments in which primates are trained to use sign languages, furthermore, are conducted under such stringently controlled situations that we learn very little about primates in the wild. 2.2.2. the vocal apparatus here is where linguists most require the expertise of biologists, those who can detail the physiology of sound production and perception. given the human laryngeal-pharyngeal complex, glottis, tongue, soft palate, nasal passage, teeth and lips, how and why have humans–only humans–come to have vocal tracts with parts uniquely, interdependently arranged to enable speech? among those varied parts of the human vocal tract, the larynx has received the most attention. nineteenth century anatomists noted that other primates’ larynxes are not as low as those of humans. then, in the 1960s, lieberman highlighted the acoustic implications of this fact: only the human larynx is low enough (and the human pharynx relatively long enough) to produce the phonology of human language. since then, physical anthropologists have been 4 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 5 striving to determine the speech abilities of earlier hominids by estimating the length of their pharynx and the position of their larynx. soft tissues decay rapidly, so researchers must base their estimates upon speculations about the position of the hyoid bone, from which the larynx is suspended. so far, such questionable data “do not appear to provide reliable indicators of the speech abilities of extinct hominids” (fitch 2000:262). because differences between larynxes entail trade-offs, it is not clear what these facts about the larynx mean to evolutionary theory. with longer supralaryngeal vocal tracts (and hence lower formant frequencies), it has been assumed that lower larynxes enable animals to sound larger, and therefore too formidable to attack. this is the ‘size exaggeration hypothesis’ (fitch 2000:264, lieberman 2003:258), which some claim gave humans a selective advantage over other primates. but with their higher larynxes, other primates can swallow liquids or solids while breathing; humans cannot. some believe this offers non-human primates significant survival advantages, but evolutionary biologists disagree about whether it would be a greater advantage than that possible advantage provided by lower formants, or even by speech. not all of the attention paid to the vocal apparatus has been focused upon the larynx. lieberman (2003:258-62) and fitch (2000:264-5) mention the tongue and lips in regard to how we use them to form some vowels, which apes cannot produce. this does not seem central to the discussion of survival or other evolutionary selection pressures. iain davidson, finally, includes a table which summarizes several theories about connections between language evolution and archaeological measurements of skeletal indicators for hominid brains, spinal cords, hypoglossal canals, hyoids and vocal passages (2003:145). while those findings remain inconclusive, they represent interesting possibilities for much more future research. 2.2.3. cognition and logic the size and shape of brains can be estimated from the crania of fossilized skulls, but minds and thoughts leave no fossil evidence. einstein, among many, commented often about how little we understand about the physical brain’s relation to the mind: we have the habit of combining certain concepts and conceptual relations (propositions) so definitely with certain sense experiences that we do not become conscious of the logically unbridgeable gulf which separates the world of sensory experiences from the world of concepts and propositions (1944:287). decades later, andrew huxley, president of england’s royal society, complained that neo-darwinists have “too often swept under the carpet the biggest problem in 5 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 6 biology, the existence of consciousness” (1981:v). there may be little to say about what earlier hominids had in mind, but davidson (2003:140-57) strives to employ archaeological evidence of tool-making to indicate their mental capabilities. tool-making and language are both products of intelligence, but his leap from chipping stone to language does not clearly differentiate between physical dexterity and that of mind. jackendoff, pinker, dunbar and others also write about the cognitive and logical essentials of language, what a mind must be able to do in order to use language as we do. jackendoff’s discussions of the relationship between cognitive conceptual structures and language in foundations of language appear perceptive and bear extensive examination. one of his claims: conceptual structure is not part of language per se–it is part of thought. it is the locus for the understanding of linguistic utterances in context, incorporating pragmatic considerations and “world knowledge”; it is the cognitive structure in terms of which reasoning and planning take place. that is, the hypothesized level of conceptual structure is intended as a theoretical counterpart of what common sense calls “meaning” (2002:123). our brains contain what he calls the f-mind (‘functional mind,’ much like our everyday use of ‘mind’), which contains conceptual structures, our cognitive organization of what we know and think. conceptual structures are connected to what happens in the real world through cognitively constructed percepts. percepts are the bundled chunks of experience (esp. sights and sounds) or perceptions delivered to the conceptual structures in the mind (or brain) by our “perceptual systems [which] evolved in order that organisms may act reliably in the real world” (2002:307). that which our conceptual structures perceive is not reality itself (not the events themselves), but it is “reality for us” (2002:309)–good enough to enable us to survive and function in the world. our brains, in other words, do not ‘get’ the real world directly; rather, they indirectly receive perceptions of the entities and experiences of the real world: events in world > perceptual systems > percepts > conceptual structures in f-mind jackendoff emphasizes that each stage in this cognitive process is physical, that the sequence from external event to our meaning of it is (and can only be studied as) a natural, material event: people find sentences (and other entities) meaningful because of something going on in their brains... there is no magic. that is, we seek a thoroughly naturalistic explanation that 6 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 7 ultimately can be embedded in our understanding of the physical world (2002:268). jackendoff then has quite detailed diagrams of relationships between meaning (those conceptual structures in our minds) and the grammatical functions of language. but, of course, between concepts and grammar are the words, the linguistic symbols employed to index the elements of a concept. jackendoff and others discuss this at length. for language to happen, two parties must “have the cognitive components that allow meaning to be attached to arbitrary signals in order to transfer information from one mind to another” (dunbar 2003:225). those parties must have a theory of mind: they must realize the existence of other consciousnesses, that they can affect the attention and intent of other minds. so far, it does not appear that any non-human primates have this capacity (tomasello 2003:100-1). animals do seek to affect each other’s behavior, but not, from what we can tell, to affect each other’s minds through information transfer. what dunbar calls ‘arbitrary signals’ are usually called linguistic symbols. like dunbar, jackendoff (1999:273), deacon (2003:117-9), pinker (2003:17) and others emphasize that symbols must be arbitrary: there is, for example, no iconic resemblance of any kind between an elephant and the word (the symbol) used to index it in spoken language. while most written symbols are also arbitrary, there remains in some languages some residual iconic or visual resemblance between a symbol and that which it represents. in chinese, several characters (e.g. those for mountain and door) still retain some faint resemblance to that which they index. symbols are triadic conventions involving the speaker, the hearer, and a referent; because they are arbitrary, they work only if agreed upon by those who employ them. agreement entails a theory of mind, of course, and must apparently be reached through the use of language (i.e. other symbols). symbols cannot be conventionalized by using only icons and indices; symbols require pre-existing symbols (oller 2002:17). deacon does not address this problem in his discussion of the logic of icons, indices and symbols (2003:111-39, 1997:70ff). and if, as jackendoff claims, “symbol use [is] the most fundamental factor in language evolution” (1999:273), then oller’s point above reveals a serious question to consider in the evolution of language from non-language. oller is not the first logician to argue for symbols’ dependence upon symbols, for, even a century ago, charles sanders peirce did so in his 1902 paper “the icon, index and symbol” (156-173). both deacon and oller frequently refer to peirce’s writings. so if vervet monkeys warn each other with one sound when they see a leopard, another when they see a snake, and a third when they see an eagle, are they using symbols? no, for symbols are not situation-specific; they are used to index something not present, an activity only humans are able to do. while it appears that “some of the foundations of the human conceptual system are present in other primates, such as the major subsystems dealing with spatial, causal, and social reasoning” (pinker and jackendoff 2005:205), those primates appear able 7 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 8 neither to know nor to index what they do not immediately perceive with their senses. of course, humans go far beyond merely using an arbitrary symbol to stand for some object or event–we can combine a finite range of sounds into a seemingly infinite range of segment-symbols (syllables, words, phrases, clauses, etc.) in such a way that we can utter propositions about complex relations among objects and events. just as our minds deal in reference, predication, categorization and the like, so does our language. and that is syntax. 2.2.4. syntax when noam chomsky began writing about the role of language in issues of nature and nurture, he may have been a significant catalyst for the resurgence in discussion of the evolution of language. after proposing almost half a century ago that all humans are born with some innate biological capacity for syntax (universal grammar, or ug), chomsky and others have spent much of the decades since revising theories of what it is and where it came from. chomsky has consistently denied that ug and other components of the human language capacity can be the result of natural selection acting upon random variation (1975:59, 1978:38-9). pinker and bloom (1990:707), however, argued that the ability to use syntax is an example of adaptive complexity, and that it certainly must have evolved by conventional darwinian means, for natural selection is the only means known to science which can produce such adaptive complexity. thirteen years later, pinker still uses the same argument, a form of questionbegging: we have syntax, so natural selection must have done it, for only natural selection can do things like that (2003:21-2). this is not too far removed from “the bible is true, because god said so in the bible.” deacon (1997:258) also uses the same logic to reach the same conclusion. deacon, pinker and bloom are not alone in this position. dawkins (1976) and dennett (1995), neither of them linguists, are also adaptationists, using the same logic to claim that syntax emerged as a biological adaptation to the environment. jackendoff, too, has joined this camp (1999:272, 2002:231-5). arguing from anti-adaptationist perspectives are almost everyone else in the discussion: chomsky, gould, bickerton, newmeyer, kirby, hurford and others. basically, the anti-adaptationist position says that there is no selective advantage (no increased survival fitness) to our complex conceptual structures or to the expressions thereof (e.g. hurford 2003:44-9). chomsky, especially, opposed any suggestion that the “principles of ug arose by virtue of their utility in fostering the survival and reproductive possibilities of the individuals possessing them” (newmeyer 2003:60). he and some others in the camp have allowed for the possibility of an ‘exaptationist’ scenario, in which ug may have arisen as a byproduct of other evolutionary processes. still, he insisted that “it would be a serious error to suppose that all properties, or the interesting properties of structures that evolved, can be ‘explained’ in terms of natural selection” 8 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 9 (chomsky 1975:59). anti-evolutionists agree, and some evolutionary biologists have accused chomsky of closet creationism. jackendoff calls this a “retreat to mysticism” (2002:234). among evolutionary linguists, then, a lot of attention has been given to how some syntactic properties resemble (and may therefore be an evolutionary by-product of) some physical survival ability. syntactic recursion, for example, might be said to parallel an animal’s ability to seek and find food inside an egg inside a nest inside a hole inside an old tree inside a forest. recursion itself has recently received much attention by some in this field. bickerton says, “syntax forms a crucial part, arguably the most crucial part–since no other species is capable of it–of human language. if we are going to explain how language evolved, we have to explain how syntax evolved” (2003:87). but hauser, chomsky and fitch claim that it is only recursion, rather than all of syntax, that is uniquely human (2002:1570). in an important 2002 paper in science, hauser, chomsky and fitch speak of a‘faculty of language in the broad sense’(flb), which includes both the needed sensory-motor abilities (vocalization, breathing, hearing, vision, and gestures) and the conceptual-intentional abilities (cognitive grasp of reference, predication, etc.). these two groups of abilities, and others, according to hauser, chomsky and fitch, evolved as adaptations to an environment, or as by-products of such adaptations. within this shared flb is what they call the ‘faculty of language in the narrow sense’ (fln), consisting only of recursion. this faculty is not shared with other species; it is recently evolved and unique to humans (hauser, chomsky and fitch 2002:1573). while their paper makes other points, the ‘recursion only’ claim is its primary one. pinker and jackendoff respond in a lengthy article in cognition, the main idea of which is that, while recursion is uniquely human, there are other key elements of language which are also unique to humans, among them “phonology, morphology, case, agreement, and many properties of words” (2005:201). they also maintain that “language is a complex adaptation for communication which evolved piecemeal...” (201). in his 1999 article, and in his 2002 book, jackendoff proposes some possibilities for how the evolution of syntax may have proceeded from context-dependent single symbol (word) utterances through the concatenation of words to more fully-formed syntax (1999:272-9, 2002:242-64). as ‘fossilized’ evidence of the transitional stage (mere concatenation), he offers some english compounds of differing relations between their parts: doghouse, housedog, snowman, man-eating, garbage man, etc. (1999:276, 2002:249-50). far more has been written about the possible evolution of syntax. close attention, though, should be paid to the possibly crucial role of symbolic logic and semiotics in understanding the evolution of syntax. deacon, jackendoff and tomasello have avoided, or only barely touched upon, the serious problem raised by peirce and oller above. the logical and mathematical prerequisites of human communication (as defined by information theory, which began with claude shannon’s 1949 mathematical theory of communication) would also seem to apply, but that shall have to be considered in another paper. 9 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 10 2.2.5. social elements according to bickerton (2003:82), “the most crucial thing to grasp about the emergence of symbolic representation is that it must have been primarily a cultural rather than a biological event.” how communities without symbols cooperated to conventionalize the use of symbols is not explicit in his account, but he appears to make indirect reference to this gap: it is no accident that in most, if not all, computer simulations of language evolution, the self-organizing ‘agents’ already know what their interlocutor means to say. if the problem space were not limited in this way, the simulations simply wouldn’t work–the agents would never converge on a workable system. but such unrealistic initial conditions are unlikely to have applied to our remote ancestors (2003:86, italics in original). in addition to this problem, of course, is perhaps an even greater one for evolutionary linguists. biological evolution is a tale of competition, of natural selection eliminating the less fit and empowering the more fit. if this is the case, how could the evolution of language have occurred, if language (even the mere agreement upon symbols) requires the cooperation of hominids competing with each other for survival? darkness and tall grass may have caused gestural communication to give way to more socially beneficial vocal communication, but how does this reconcile with ‘survival of the fittest?’ it is to each hominid’s survival advantage that his/her competitors not know his/her intent. numerous studies by tomasello, hare, and call, for example, have revealed how animals (especially apes, goats, and dogs) strive to follow each other’s gaze when competing for food. such observations about cooperation and competition are not naïve responses to a merely apparent contradiction, or evolutionary biologists and evolutionary linguists would not struggle so vigorously to counter it. knight, studdert-kennedy and hurford do so creatively: ...language is no ordinary adaptation, but will require ‘special’ darwinian explanation, ...which isolates biologically anomalous levels of social cooperation as central to the evolutionary emergence of language... language, in short, is remarkable–as will be any adequate darwinian explanation of its evolution (2000:12). richard dawkins has put forth an extremely creative ‘special darwinian’ theory to address this anomaly, his selfish gene theory. very simply, the theory asserts that within each organism is a selfish (or selfishness) gene, bent on survival, and willing to put up with temporary inconveniences such as altruism, co-operation, even sacrifice, in order to achieve longer term viability. this gene 10 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 11 ‘lives on’ across many generations, and it can even foresee that short-term anomalies benefit long-term patterns (dawkins 1976:2-3). if such a gene has no consciousness, it is hard (perhaps impossible) to grasp how it might ‘foresee’ anything; if it has consciousness, believing in it differs little from faith in a god. dawkins’ approach is one of a variety of attempts to address the issue of depending upon a theory about competition to explain the development of an intrinsically co-operative activity. others have written much more about the integral part played by a theory of mind, by shared attention, by the triadic use of symbols, by culture, and so on. those, however, are also beyond the scope of this very brief introduction to the field. in a summarizing statement, jackendoff does use evolution as a metaphor when he says, “languages may change and ‘evolve’ in the sense of cultural evolution, but as far as can be determined, this is in the context of a fully biologically evolved capacity” (2002:232, italics original). each of the scientists above makes frequent reference to tenets of physical evolution, for it is upon a foundation of evolutionary biology that evolutionary linguistics is building. and some want even more: chomsky has stressed that language is a biological phenomenon. but prevalent contemporary brands of linguistics neglect the evolutionary dimension. the present facts of language can be understood more completely by adopting an evolutionary linguistics, whose subject matter sits at the end of a long series of evolutionary transitions, most of which have traditionally been the domain of biology... the key to explaining the present complex phenomena of human language lies in understanding how they could have evolved from less complex phenomena... modern languages are learned by, stored in, and processed online by evolved brains, given voice by evolved vocal tracts, in evolved social groups (hurford 2003:40). newmeyer also believes in recruiting other scientists from other fields to the study of language evolution, and he states this strongly: . . . if the properties of universal grammar are what they are as a result of physical principles, then it falls to the physicist and molecular biologist to unravel language origins, not to the theoretical linguist (2003:60). obviously, then, given the interdependencies of speech organs and phonology, of the brain and meaning, linguistics and biology are necessarily and irrevocably entangled. 11 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 12 if neo-darwinian evolutionary biology is supported by solid facts and logic, then linguistics can take this foundation for granted and focus instead upon how best to construct an account of the emergence of human speech. but is it possible that linguists are assuming too much? now, with some sense of the current state of evolutionary linguistics, let us briefly consider the evolutionary biology upon which it apparently depends. 3. evolutionary biology the details of what is almost universally taught about evolutionary biology are too well known to describe at length. over billions of years, from non-organic matter (often called the ‘pre-biotic soup’) organic compounds emerged by sheer chance. from this organic matter, combined with time and chance, the first living, self-replicating cell appeared by some process of self-organization. and from that first cell evolved many more single-celled, then multi-celled, ever more complex organisms: bacteria, amoebae, invertebrates, vertebrates, fish, amphibians, reptiles, birds, mammals, primates, and ultimately, linguists. the genius of darwin was in proposing that the driving force for progress (or the filter which preserved the superior and eliminated the inferior) was natural selection. when darwin’s theory was informed by the discovery and application of more modern sciences (especially genetics), the result was called neo-darwinism, now the dominant theory in evolutionary biology. most renowned scientists in the field (cousteau, dawkins, dennett, dobzhansky, gould, haldane, huxley, leakey, mayr, sagan, et al) are or were neo-darwinists. according to neo-darwinists, then, as living organisms evolved at micro (genetic) and macro (species) levels, those variations or mutations which provided survival fitness were selected and passed on by those who had them. variations which provided no survival fitness, no selective advantage, were eliminated as those who possessed them died off. hence, the popular expression ‘survival of the fittest.’ this, roughly, is neo-darwinian microand macroevolution. 3.1. established fact? if one reads the popular science writing of carl sagan, stephen jay gould, richard dawkins, and daniel dennett, if one peruses the pages of national geographic, if one wanders the websites of the national center for science education (ncseweb.org) or the american association for the advancement of science (aaas.org), one will likely conclude that neo-darwinian progressive macroevolution is solid, unquestionable fact, with a few stray details still left to be filled in. some (mostly religious) people are not yet totally convinced, but it is only a matter of time before everyone is enlightened. dennett states this even more strongly, seeming to prefer ad hominem to argument: 12 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 13 to put it bluntly but fairly, anyone today who doubts that the variety of life on this planet was produced by a process of evolution is simply ignorant–inexcusably ignorant–in a world where three out of four people have learned to read and write. (dennett 1995:46) the writings of dennett, dawkins, gould, sagan and others are featured in skeptic magazine (www.skeptic.com). differing views are often caricatured there, dismissed as those of “hordes of creationists that infest the american pseudointellectual landscape and stubbornly try to legislate scientific ignorance in our public schools” (pigliucci 2001:54). if scientifically-challenged school board members are neo-darwinism’s only opposition, then evolutionary linguists likely need not concern themselves with such disputes. but this is not the case. some of neo-darwinism’s recent, high-profile challenges have come from those who hold to the theory of intelligent design. as something of a philosophy of science ‘think tank,’ i.d.’s people do not conduct laboratory experiments in pursuit of hard data to support a competing theory; rather, they apply accepted principles of science, math and logic to highlight areas in which neo-darwinism has more work to do before it can claim to represent unassailable, demonstrable fact. while mainstream media claim or imply that i.d. people are fundamentalist christian creationists, little research is needed to learn that numerous prominent adherents do not fit that description. mustafa akyol, michael behe, gertrude himmelfarb, seyyed hossein nasr, gerald schroeder, and vladimir voeikov may be amused or troubled by such a simplistic caricature, by being dismissed with little more than a transparent ad hominem. also contrary to most media, i.d.’s primary unifying focus is neither religion nor public school curricula (though courts have rejected any discussion of i.d. in public schools). i.d. argues that, while the explanatory force of natural selection remains great, it is still unable to account for much of the specified complexity observable in the universe. though no conflict exists between random variation and accidental complexity, specified (or functional) complexity usually indicates intelligence. the search for extra-terrestrial intelligence (or seti), for example, is predicated upon precisely that fact. if seti were someday to detect radio pulses in morse code, or sequences of prime numbers, it will have detected specified complexity (or even language), and would assume that some purposive intelligence is ‘out there.’ few scientists attack seti, though, and the difference, of course, lies in the respective intelligences that i.d. and seti seek. among i.d.’s apparent leaders (michael behe, william dembski, stephen meyer, and others) are credentialed, practicing, peer-reviewed, published scientists, or scholars in non-scientific fields, not exactly dennett’s “inexcusably ignorant” folk. their writings detail technical, procedural, or logical/conceptual weaknesses in neo-darwinism. here, for example, is just one of many questions dembski raises, one clearly relevant to the evolution of language: 13 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 14 ...there is an inherent tendency in evolving systems for selection pressures to force such systems toward simplicity. this is not to say that darwinism requires or entails that evolution proceed toward simplicity. the point is simply that darwinism, in itself, does not mandate increasing complexity and inherently favors simplicity. thus, if we see increasing complexity, something besides darwinism must be at work (dembski 2004:256). by implication, perhaps, entropy, information theory, complexity theory, and occam’s razor might all be productively applied to questions about the relationship between simplicity and complexity in the evolutionary emergence of language. dembski is referring to ongoing discussion among scientists as varied as stephen jay gould, stuart kauffman, and hubert yockey. for those troubled by the apparently metaphysical implications of i.d., the skepticism of various credentialed, non-i.d. scholars certainly warrant attention. franklin harold, for example, is emeritus professor of biochemistry and molecular biology at colorado state university. from his oxford university press text, the way of the cell: life arose here on earth from inanimate matter, by some kind of evolutionary process, about four billion years ago. this is not a statement of demonstrable fact, but an assumption almost universally shared by specialists as well as scientists in general. it is not supported by any direct evidence, nor is it likely to be. ...the reasons for the general consensus are, first, the lack of a more palatable alternative; and second, that absent the presumption of a terrestrial and natural genesis there is no basis for scientific inquiry into the origin of life (harold 2001:236-7). recall that newmeyer (2003:60), above, recommends that linguists turn to molecular biologists (such as harold) to secure answers to persistent questions about the evolution of life and language. another skeptic, a significant figure in chemistry, genetics and microbiology, is the university of chicago’s robert shapiro. he points out that there are far more unresolved questions than answers about evolutionary processes, and contemporary science continues to provide us with new conceptual possibilities. unfortunately, readers may remain unaware of this intellectual ferment because… serious open-minded discussions of the impact of discoveries in molecular biology are all too rare. the possibility of nondarwinian scientific viewpoints is virtually never considered (shapiro 1998). 14 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 15 shapiro is one of dawkins’ professional peers who finds both neo-darwinian and creationist accounts scientifically unsatisfactory, and who wants to reopen discussion. this ‘intellectual ferment,’ as he puts it, is no threat; it is, rather, essential for the progress of scientific knowledge. later, he writes: ...our current knowledge of genetic change is fundamentally at variance with postulates held by neo-darwinists... nonetheless, neo-darwinist writers like dawkins continue to ignore or trivialize the new knowledge... this is to be expected from creationists, who naturally refuse to recognize science’s remarkable record... but the neo-darwinian advocates claim to be scientists, and we can legitimately expect of them a more open spirit of inquiry. instead, they assume a defensive posture of outraged orthodoxy and assert an unassailable claim to truth, which only serves to validate the creationists’ criticism that darwinism has become more of a faith than a science (shapiro 1998). although shapiro’s scientific achievements and credentials are remarkable, with uncompromisingly forthright comments such as these, he may be running the risk of being professionally shunned. mary midgley, philosopher of science at the university of newcastleupon-tyne, has two fascinating titles, evolution as religion (1985) and science as salvation (1992), both of which list many examples of the phenomena shapiro describes above. but because midgley is not a lab coat scientist, her observations are perhaps more easily discounted by the likes of dawkins and dennett. stephen jay gould (until his death harvard’s renowned evolutionary paleobiologist) figures prominently in skeptic, and in the writings of other neodarwinists. he echoes the concerns of shapiro and midgley: “...we have persecuted dissenters, resorted to catechism, and tried to extend our authority to spheres where it has no force...” (1977:146). this is precisely what concerns shapiro, who also says, dogmas and taboos may be suitable for religion, but they have no place in science. no theory or viewpoint should ever become sacrosanct, for experience tells us that even the most elegant laws of nature ultimately succumb to the inexorable progress of scientific thinking and technological innovation (shapiro 1998). if neo-darwinism is established fact, then it has nothing to hide, for as john milton says in areopagitica, “who ever knew truth put to the worse, in a free and open encounter?” (1674:746). science need not fear i.d., harold, shapiro or midgley, or others like them who agree with much of what scientists say when they write of scientifically demonstrable facts. 15 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 16 numerous scientists agree with evidence for microevolution within species but question the macroevolutionary (producing new species) claims of neodarwinism. opponents with purely religious motives (who offer no scientific evidence or logical argument) need not concern scientists. if connections were to exist between life’s origins and a biblical account of creation, or between the diversity of languages and a biblical story of babel, such connections would likely be inaccessible to the scientific methods we now employ. significant dissent, however, has arisen not just from these corners but among scientists who are content to leave bible stories to the clergy. clearly, skepticism about neo-darwinism’s claims and presuppositions is not just emerging from intelligent design, or from a few isolated religious institutions or rural school districts. the wide variety of scholars above are just several of many calling for more transparent discussions of macroevolution’s unsettled questions. in the halls of scientific academia, some of neo-darwinism’s most important assumptions still warrant truly objective (re)consideration. it does not yet appear that the neo-darwinian approach to biology is in imminent danger of collapse; nevertheless, the foundations of evolutionary linguistics are not as solid as some have apparently assumed. a consensus among scientists certainly exists about evolution, but consensus is not scientific evidence. we did not lose the flat earth or gain the periodic table by sheer numbers of voting scientists, and consensus will not serve us well if we try using it to prove or disprove anything in evolutionary linguistics. 3.2. linguistics’ dependence upon evolutionary biology as we have seen above, language is much more than merely biological; it is conceptual/logical, social, mathematical, and more. anatomy and acoustics, for example, inform linguistics’ understanding of the productions and perceptions of sound. between the sounds we share with animals, however, and the meanings to which only humans can harness them, there remains a vast and still poorly understood gulf. to grasp how humans bridged this gulf, it seems that linguistics needs more than evolutionary biology has to offer. psycholinguistics already depends upon developmental psychology and cognitive science. sociolinguistics, too, depends upon sociology and political science. these are fields formerly dominated by two of darwin’s most influential disciples, freud and marx. if ‘oedipal’ and ‘proletarian’ now sound like quaint old jargon, they may also caution linguists against depending too heavily (or even exclusively) upon one still unstable and fallible scientific theory. this will hardly limit evolutionary linguistics, for there also exist other fields whose intersections with linguistics warrant far more research. the rich writings of charles sanders peirce and claude shannon, mentioned above, are seldom referenced in linguistics, yet from them we have much to learn about problems in the logical complexities and mathematical probabilities of homo 16 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 17 sapiens sapiens acquiring the capacity to use a finite range of linguistic symbols to express a potentially infinite range of concepts. plotkin and nowak (2000) and oller (2002) are some of the few who have begun exploring how we might benefit from a synergy among the studies of linguistics, information theory, mathematical probability, complexity theory, the logic of symbols, and more. regardless of the sciences to which linguists turn, we will benefit by a commitment to responsibly follow evidence, rather than consensus, even when it leads us to temporarily inconvenient conclusions: i will not inquire as to the details of how increased expressive power came to spread through a population, nor how the genome and the morphogenesis of the brain accomplished these changes. accepted practice in evolutionary psychology… generally finds it convenient to ignore these problems; i see no need at the moment to hold myself to a higher standard than the rest of the field (jackendoff 2002:237). for the sake of brevity and clarity, perhaps, jackendoff might be forgiven for postponing the discussion of certain tangential issues. but as linguists strive to determine and describe the origins of language, we will be wise not to ignore neo-darwinism’s glaring problems as we strive to avoid the sorts of errors that shapiro and others warn against above. “if a single conclusion drawn from [general relativity] proves wrong, it must be given up; to modify it without destroying the whole seems to be impossible” (einstein 1934:60). linguists need not, indeed cannot, surrender all that we have learned from evolutionary theory, but we may regret not exercising more caution, or not asking more questions about the evolutionary biology upon which we currently depend, and not imitating einstein’s increasingly rare intellectual modesty. references arbib, michael. 2003. “the evolving mirror system: a neural basis for language readiness.” in christiansen, morten and simon kirby. language evolution. new york: oxford. 182-200. bickerton, derek. 1981. roots of language. ann arbor, mi: karoma. bickerton, derek. 2003. “symbol and structure: a comprehensive framework for language evolution” in morten christiansen and simon kirby (eds.). language evolution. new york: oxford. 77-93. chomsky, noam. 1975. reflections on language. new york: pantheon. chomsky, noam. 1978. rules and representations. new york: columbia. christiansen, morten and simon kirby (eds.). 2003. language evolution. new york: oxford. 17 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 18 corballis, michael. 2003. “from hand to mouth: the gestural origins of language” in morten christiansen and simon kirby (eds.) language evolution, 201-18. new york: oxford davidson, iain. 2003. “the archaeological evidence of language origins: states of art.” in morten christiansen and simon kirby (eds.). language evolution. new york: oxford. 140-57. dawkins, richard. 1976. the selfish gene. london: oxford. deacon, terrence. 1997. the symbolic species: the co-evolution of language and the brain. new york: norton. deacon, terrence. 2003. “universal grammar and semiotic constraints” in morten christiansen and simon kirby (eds.). language evolution, 111-39. new york: oxford. dembski, william. 2004. the design revolution. downers grove, il: intervarsity. dennett, daniel c. 1995. darwin's dangerous idea: evolution and the meanings of life. london: penguin. dunbar, robin. 2003. “the origin and subsequent evolution of language.” in morten christiansen and simon kirby (eds.). language evolution, 219-34. new york: oxford. einstein, albert. 1934. the world as i see it. new york: covici-friede. einstein, albert. 1944. “remarks on russell’s theory of knowledge.” in paul arthur schilpp (ed.), the philosophy of bertrand russell. new york: tudor. fitch, w. tecumseh. 2000. “the evolution of speech: a comparative review.” trends in cognitive science: 258-67. gould, stephen jay. 1977. ever since darwin. new york: norton. hare, brian and michael tomasello. 1999. “domestic dogs use human and conspecific social cues to locate hidden food.” journal of comparative psychology 113: 173-7. harold, franklin m. 2001. the way of the cell. new york: oxford university press. hauser, marc, noam chomsky and w. tecumseh fitch. 2002. “the faculty of language: what is it, who has it, and how did it evolve?” science 298: 1569-79. hewes, g. w. 1973. “primate communication and the gestural origin of language.” current anthropology 14: 5-24. hurford, james r. 2000. “introduction: the emergence of syntax” in chris knight, michael studdert-kennedy and james hurford. the evolutionary emergence of language, 219-30. new york: cambridge. hurford, james r. 2003. “the language mosaic and its evolution.” in morten christiansen and simon kirby (eds.) language evolution, 38-57. new york: oxford. huxley, andrew. 1981. supplement to royal society news 12: v. jackendoff, ray. 1999. “possible stages in the evolution of the language capacity.” trends in cognitive science 3: 272-9. 18 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 evolution and language 19 jackendoff, ray. 2002. foundations of language: brain, meaning, grammar, evolution. new york: oxford. knight, chris, michael studdert-kennedy and james r. hurford. 2000. the evolutionary emergence of language. new york: cambridge. lieberman, philip. 2003. “motor control, speech and the evolution of human language.” in morten christiansen and simon kirby (eds.). language evolution, 255-71. new york: oxford. lyons, john. 1970. new horizons in linguistics. baltimore, md: penguin. lyons, john. 1977. semantics. new york: cambridge. macneilage, peter and barbara davis. 2000. “evolution of speech: the relation between ontogeny and phylogeny.” in chris knight, michael studdertkennedy and james r. hurford. the evolutionary emergence of language, 146-60. new york: cambridge. midgley, mary. 1985. evolution as a religion. london: methuen. midgley, mary. 1992. science as salvation. london: routledge. milton, john. 1674. “areopagitica.” in merrit hughes (ed.). 1957. john milton: complete poems and major prose, 746. indianapolis, in: bobbs-merrill. newmeyer, frederick. 2003. “what can the field of linguistics tell us about the origins of language?” in morten christiansen and simon kirby (eds.). languageevolution, 58-76. new york: oxford. oller, john, jr. 2002. “languages and genes: can they be built up through random change and natural selection?” journal of psychology and theology 30: 2640. peirce, charles sanders. 1932 [1902]. “the icon, index, and symbol,” in c. hartshorne and p. weiss (eds.) collected papers of c. s. peirce, vol. 2: 156173. cambridge, ma: harvard. pigliucci, massimo. 2001. “review of mathematics and evolution, by sir fred hoyle.” skeptic 8(4): 54. pinker, steven. 2003. “language as an adaptation to the cognitive niche.” in morten christiansen and simon kirby (eds.). language evolution, 16-37. new york: oxford. pinker, steven and p. bloom. 1990. “natural language and natural selection.” behavioral and brain sciences 13: 707-784. pinker, steven and ray jackendoff. 2005. “the faculty of language: what’s special about it?” cognition, 95: 201-36. plotkin, joshua and martin nowak. 2000. “language evolution and information theory.” journal of theoretical biology 205: 147-59. rizzolatti, giacomo and michael arbib. 1998. language within our grasp. trends in neurosciences 21(5): 188-194. shannon, claude and warren weaver. 1949. the mathematical theory of information. urbana, il: university of illinois press. shapiro, james a. 1998. scientific alternatives to darwinism: is there a role for cellular information processing in evolution? retrieved from world wide web: http://www.asa3.org/archive/asa/199809/0015.html 19 stebbins: the evolution of evolutionary linguistics published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 20 tomasello, michael. 2003. “on the different origins of symbols and grammar.” in morten christiansen and simon kirby (eds.). language evolution, 94-110. new york: oxford. tomasello, michael, josep call and brian hare. 1998. “five primate species follow the visual gaze of conspecifics.” animal behaviour 55: 1063-9. yockey, hubert. 2005. information theory, evolution and the origin of life. new york: cambridge. 20 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/1 doi: https://doi.org/10.25810/vefc-nw72 colorado research in linguistics 6-2007 the evolution of evolutionary linguistics jeff r. stebbins recommended citation untitled language policy colorado research in linguistics. june 2006. vol.19, issue 1. boulder: university of colorado. © 2006 by michael f. thomas. book review bernard spolsky. language policy. cambridge: cambridge university press. 2004. 262 pages. isbn-13: 9780521011754 / isbn-10: 0521011752. $29.99 us. reviewed by michael f. thomas when first approaching a field of study as diverse as 'language policy', it’s easy to become disoriented under the avalanche of facts and patterns--educational policies, distinct languages spoken in a region, historical factors, legal issues, linguistic ideologies, nationalisms--and then become hard pressed to see how such divergent forces interact. in his book language policy, spolsky seeks to mediate this overload of information by providing a model to organize it. the basic premise of the model is that language 'policy' is best understood as the relationship between three factors; ideology, management and practice. management is the explicit attempt of a locus of power (such as the state) to manipulate language practices and ideologies. practice simply means how linguistic resources are habitually utilized in a speech community. ideology is the system of beliefs about language varieties and linguistic choices. an example of why this tripartite division is important can be seen in the three-language formula in india, which states that all indians should know the national language, hindi, the regional language of the state and their mother tongue. in the realm of management, children have a right to mother tongue education. the law reflects the dominant national ideology of valuing all languages in india. however, in practice, relatively few children receive instruction in their mother tongue. this is due to a number of factors, including limited resources for the publishing of educational materials in all of the languages of india, unclear distinctions between languages as in the case of dialect chains, local ideologies differing from national ideologies, etc. when analyzing the language policy of india one must look beyond the law and see if the ideology upon which the law is based is actually reflected in the practices of the various speech communities of the nationstate. in addition to this three-way distinction for analyzing language policy, spolsky asserts that three other assumptions are also necessary. first, language policy is not only concerned with named varieties. as illustrated in the example of india, the naming of varieties is in itself sometimes politically motivated (e.g. which variety in a dialect chain gets afforded official status as the standard?) furthermore, other varieties may have unofficial implications in their use. in the us, decisions made regarding the use of aave in the schools are within the realm of language policy whether they are made on an ad hoc basis by individual instructors, as in some areas, or governed by local administrative policy in others. second, language policy must be understood as operating within a speech community. ideologies relevant to the same variety of a given language will differ depending on the speech community where it is being used. third, language policy must be understood as functioning within a complex ecological relationship between linguistic and non-linguistic factors. often the non-linguistic factors include access to resources; as such, language policy often serves as a surrogate political issue for other ideological agendas. chapter 1 gives an overview of these three factors--management, practice and ideology--followed by multiple examples in chapters 2 and 3. the examples are loosely organized around themes in the regulation of language, such as 'driving out the bad' (chapter 2), and 'pursuing the good and dealing with the new' (chapter 3). spolsky’s theory of 1 thomas: language policy published by cu scholar, 2006 colorado research in linguistics, volume 19 (2006) 2 language policy is laid out in chapter four, 'the nature of language policy and its domains.' the remainder of the book discusses language policy as it applies to the modern nation-state. chapters 5 through 10 focus on monolingual polities. chapter 5 focuses specifically on france and iceland. then, in chapter 6, spolsky moves on to discuss the current spread of english around the globe, including the implications for any nation-state with an explicit monolingual policy and various reactions to this spread. chapter 7 discusses the problems of analyzing us language policy. here, the major points touch on the fact that education policy is set at the state rather than federal level, the hegemony of english as a national although not official language, and the fact that language policy issues are generally approached as civil rights issues under title vi of the civil rights act. chapter 8 discusses the various approaches to language rights that have been taken over the centuries in western societies. chapter 9 discusses post-colonial countries and the relationship between minority language users and the official and national languages of the countries the official languages generally being the languages of the colonizers. chapter 10 discusses monolingual polities with recognized linguistic minority groups. chapter 11 begins a discussion of multilingual polities and the problems faced by policy analysts and implementers who hope to partition the linguistic space. chapter 12 discusses attempts to resist language shift in the context of preserving indigenous languages as well as returning to the discussion of maintaining the standard brought up in chapter 3. spolsky wraps up with a review of the model in chapter 13. spolsky's language policy contains a profusion of data from a broad spectrum of times and places. every point about policy is discussed in terms of examples. while such thoroughness is commendable and extremely useful for those doing in-depth analysis of policy issues, at times it makes it difficult for the reader to remain focused on the thread of his arguments. with so many examples being offered and the overlapping nature of the different aspects of language policy being discussed, a clearer organizational scheme would have been helpful. this book was used as the textbook for an undergraduate course on world language policies. while the number and variety of examples certainly did much to bring the topic to life, they sometimes obscured the very model which was being put forward to clarify the issues. if the book were more clearly organized along the lines of the model it contains, it would better serve as a course text. there was also one notable absence in the book. very little was said about language policy in israel, which is very surprising given israel’s uniquely successful policy of revitalizing hebrew and spolsky's own long-standing contributions to the study of that policy. that being said, the book remains an outstanding reference work for anyone wishing to become better informed on language policy issues. the model put forward by spolsky is likely to serve as a basis for much future research on the subject and the many case studies cited are a testament to the multi-faceted nature of the issues involved in dissecting language policy. michael f. thomas university of colorado department of linguistics 2 colorado research in linguistics, vol. 19 [2006] https://scholar.colorado.edu/cril/vol19/iss1/3 doi: https://doi.org/10.25810/gsa5-d942 colorado research in linguistics 6-2006 language policy michael f. thomas recommended citation microsoft word spolsky review for cril edits.doc microsoft word moeller-cril2021-proof_final.docx 1 computational morphology for language description and documentation sarah moeller university of colorado boulder while the field of linguistics has slowly but surely widened the world’s knowledge about human language, thanks in part to the recent emphasis on language documentation and description of underdocumented languages, computational linguistics has barely expanded beyond a handful of economically or politically powerful languages. this paper is a synthesis of natural language processing (nlp) models and methods and a history about how the models and methods have been applied to the study of morphological structure, particularly in low-resource languages (lrl). the paper assumes that the study of morphology has an important role to play in both nlp and linguistics. it explores the potential for discovering newer and more efficient methods while training computational morphological models on data produced during language documentation and description (ldd) field projects.1 keywords: language documentation, natural language processing, nlp, low-resource languages, machine learning 1. introduction morphology comprises word-building properties in human languages and their accompanying (morpho-)syntactic phenomena. historically, computational linguists and “paper-and-pencil linguists” have taken different and sometimes seemingly incompatible approaches to morphology (karttunen & beesley 2005). yet, despite their out-of-sync approaches, both computational linguistics and “traditional” linguistics benefit from morphological analysis (cotterell et al. 2015). for natural language processing (nlp), work with low-resource languages (lrl) is still largely uncharted territory. this paper explores the limited work in morphology by asking this question: “what [computational] methods...can detect [morphological] structure in small, noisy data sets, while being directly applicable to a wide variety of languages?” (bird 2009). the paper is organized as follows. section 2 describes the workflow and activities of ldd. section 3 sketches the history of nlp work with lrl. section 4 defines morphological analysis and looks specifically at nlp applied to morpheme segmentation and glossing and section 5 looks at the application to learning morphological inflectional paradigmatic patterns. colorado research in linguistics, volume 25 (2021) 2 2. language documentation and description in linguistics, morphological description of a broad range of languages is a foundational step towards any reasonable linguistic theory. a focus of documenting and describing underdocumented languages, including their morphological structure, has been emphasized since the 1990’s along with the development of language documentation as a distinct subfield. himmelmann (1998) defines language documentation as “a comprehensive and representative sample of communicative events [that are] as natural as possible.” woodbury (2003) defines it similarly as “comprehensive and transparent records supporting wide ranging scientific investigations of the language.” language description can be defined as work that analyzes language documentation to create “systematic presentations of the phonology, morphology, syntax, and semantics of the language” (bird & chiang 2012). the emphasis on endangered languages over the past three decades has established best practices documenting a new language (bowern 2008; czaykowskahiggins 2009; lupke 2010; vallejos 2014; rice & thunder 2017). however, the specific activities that divide the two subfields are not rigid. therefore, the current work generally refers to them together as “language documentation and description (ldd)” or “documentary and descriptive linguistics”. the workflow of ldd is not standardized, although most projects seem to follow a similar sequence. a version of one common sequence (bird & chiang 2012) is given below (the numbers are used to refer to each task, e.g., “task 2a” refers to transcription). each subsequent task progressively encompasses more description than documentation, except archiving, which comes strictly under language documentation but is logically a last step. (1) collect (audio/video recordings) naturally occurring speech (2) a) transcribe and b) translate (3) perform basic morphosyntactic analysis by segmenting the morphemes and creating morphological glosses and/or a lexicon (4) elicit morphological paradigms that reveal underlying patterns (5) prepare descriptive reports that outline the language’s structure (6) archive data in a long-term repository computational morphology for language description and documentation 3 one primary output of this workflow is interlinear glossed texts (igt), a data format distinctive to linguistics (figure 1). interlinearization is the primary task after transcription. it moves the workflow beyond simple documentation but still serves as a “preprocessing step” to language description (strictly defined) (moon, erk & baldridge 2009). it comprises annotation tasks that enrich the data with analytic information added as lines under the transcribed text (task 1 in above workflow; line 1 in figure 1). the most common lines are shown in figure 1. lines can be added in any order, but translations (task 2a; line 7) morpheme boundaries and morpheme glosses (task 3; lines 2 and 3, respectively) are usually added first. doing more annotation (e.g. lines 4-6) often happens in field projects, but translation, morpheme segmentation, and morpheme glossing are usually given the highest priority. figure 1. interlinearization: interlinear glossed texts add lines of annotation to the original text. interlinearizing data uncovers the rarer and unique linguistic phenomena. interlinearization opens the door for deeper linguistic analysis and lays the foundation for reference grammars, dictionaries, and language learning materials, but interlinearization is not sufficient to create complete grammars, dictionaries, etc. one additional descriptive task is often included: the collection of morphological inflection patterns, or paradigms, for several lemmata (task 4). inflectional colorado research in linguistics, volume 25 (2021) 4 paradigms are elicited because complete paradigms are rarely found in natural language. complete paradigms are needed to infer general rules of inflection. without translations, morpheme segmentation, and glossing, the data is understandable only to someone who already speaks the language. if no speakers are left, the data is mostly inaccessible, much like egyptian hieroglyphics before the rosetta stone was discovered. a few specially designed software tools provide limited automated assistance. the two most popular are elan (auer et al. 2010) and flex (rogers 2010). examples of their interlinearization interfaces are shown in figures 2 and 3. these tools implement hand-constructed, rule-based computational morphological parsers but rule-based parsers do not generalize to new data. flex also copies morpheme boundaries and glosses onto other words if they are identical to words that were previously annotated by hand. neither tool incorporates machine learning. figure 2. user interface for interlinearization in fieldworks language explorer (flex) displaying a manipuri [mni] text computational morphology for language description and documentation 5 figure 3. user interface for interlinearization in elan showing a practice session in english 3. natural language processing (nlp) for low-resource languages though the line between documentation and description may not be clear, one thing is clear: current methods cannot easily process large amounts of data. most archived corpora are only partly annotated because funding and time constraints do not allow complete interlinearization (cox, bouliame & alam 2019). methods currently that are today used commonly in ldd rely primarily on hand annotation which is extremely inefficient. the typical strategy of annotating texts from top to bottom is non-optimal for training a supervised machine learning model (baldridge & osborne 2008; baldridge & palmer 2009; palmer 2009). since naturally occurring speech contains many repeated linguistic structures, manual annotation has been described as repetitive, monotonous, costly, and time-consuming (duong 2017; he et al. 2016). it can take anywhere from 20 to 100 hours to transcribe (task 2a) a single hour of speech (seifart et al. 2018) and it is reasonable to assume that interlinearization (tasks 2b and 3) and eliciting morphological paradigms (task 4) require significantly more time. a recent growth of nlp interest in low-resource languages (lrl) has brought machine learning models and methods that achieve good results even with ldd field data. a notable example is elpis (foley et al. 2018), an online tool that includes a user interface accessible to those with no programming background. machine translation (mt) has also been applied to documentary data, using the output of an automatic speech recognition system as input to the mt system (anastasopoulos, chiang & duong 2016; duong et al. 2016). colorado research in linguistics, volume 25 (2021) 6 the potential for machine learning to perform morphological analysis during interlinearization has been clearly demonstrated (baldridge & palmer 2009; palmer 2009; palmer et al. 2010; xia et al. 2016). for example, felt (2012) found that when a round of annotation is done automatically by a machine learning model and then corrected by the human annotators, the annotators’ accuracy is improve if the machine learning model achieves at least 60% accuracy and significantly speeds manual annotation if it achieves an accuracy of 80%. in the area of morphological paradigm learning, the annual sigmorphon and conll-sigmorphon shared tasks (cotterell et al. 2016; cotterell et al. 2017; cotterell et al. 2018; mccarthy et al. 2019; nicolai, gorman & cotterell 2020) have developed successful methods with limited training data. although nlp interest in lrl has grown noticeably in the past few years, it is not a new area of research. since the late 20th century, nlp has taken several approaches to low-resource languages that can be classified as either rule-based (i.e., finite state transducers) (e.g., cotterell et al. 2015; forsberg & hulden 2016; moeller et al. 2018; moeller et al. 2019) or machine learning models that “learn” rules from data. machine learning approaches to lrl can be classified according to whether the training data was annotated completely (supervised) (e.g., bergmanis et al. 2017; sudhakar & singh 2017; makarov, ruzsics & clematide 2017; liu et al. 2018; makarov & clematide 2018a), partially (semi-supervised) (e.g., ahlberg, forsberg & hulden 2014), or not at all (unsupervised) (e.g., moon, erk & baldridge 2009; palmer et al. 2010; kirschenbaum, wittenburg & heyer 2012; soricut & och 2015). at first glance, unsupervised and semi-supervised learning seem most promising for ldd because they do not require as much manually annotated data. supervised learning is trained on “gold standard” annotated data. however, even though supervised learning requires annotation, it needs much less data than unsupervised learning and almost always yields better results (ruokolainen et al. 2013; cotterell et al. 2015). additionally, without annotated labels, unsupervised learning can only really cluster data by the latent patterns in the data. discovering latent patterns might be quite useful for linguists when first exploring the data; for example, frequent character patterns and substrings that a model discovers could provide an initial hypothesis to the linguist about the language’s morphological structure. however, no matter how accurate an unsupervised model may be, it cannot substitute the valuable process of manually analyzing and discovering patterns in the data. detailed analysis of new data is vital for linguists because through that process the linguist becomes familiar with the data and begins to absorb an computational morphology for language description and documentation 7 intuitive knowledge of the language. nevertheless, the latent patterns discovered by unsupervised models can have many uses such as being leveraged in a semi-supervised approach. semisupervised learning combines some supervised data with a larger set of unsupervised data (kohonen, virpioja & lagus 2010; poon, cherry & toutanova 2009). this approach is suitable if available annotated data is not adequate to effectively train a supervised model and it may be ideal for ldd because having substantial amounts of unannotated data with a small amount of annotated data is a common situation. unfortunately, real applications of semi-supervised learning, specifically for computational morphology, are relatively rare, particularly with neural networks. there are exceptions, such as ahlberg et al. (2014), where semi-supervised learning was used to induce morphological paradigms in low-resource settings. until the 2010s, most machine learning models were feature-based with hand-designed features, illustrated in figure 4. the input would be a hand-designed feature function that for morphological analysis might include 1) the whole word, 2) the position of the word in the sentence, 3) surrounding words or morphemes, 4) the pos tag of the previous morpheme/word. features are assigned weights by the model during training to achieve optimal performance according to some objective function such as classification accuracy. for example, in a morpheme segmentation task where one chosen feature is the previous word and the previous word is some form of the english “to be” verb, and the target word ends in “ing”, then the model might give a high weight to the previous word so that the model pays attention to it when deciding how to segment a word ending in “ing”. the performance of feature-based models, such as conditional random fields (crf) and support vector machines (svm), relies heavily on the manual choice of features. this could be a drawback for under-described languages, because if little linguistic description is available, how does one know which features are optimal for that language? fortunately, some feature-based models have been shown to perform reasonably well using language-independent features such as length of word or placement of letter in word (ruokolainen et al. 2016; moeller & hulden 2018). colorado research in linguistics, volume 25 (2021) 8 figure 4. feature-based machine learning requires a human to identify and extract features that a feature-based classification model such as a crf uses to provide the correct output currently, neural networks models, or deep learning models, are dominating nlp (goldberg 2017). even though they outperform older, feature-based models on almost all tasks, they did not become popular until the mid-2010s because they require greater computing power and, for some tasks, train more slowly (cotterell & heigold 2017). neural networks, illustrated in figure 5, refers to a family of supervised machine learning models that are composed of layers of statistical units. the layers essentially substitute the feature engineering needed in non-neural machine learning. multiple embedded layers allow the model to look at an exponential number of “semantically” neighboring instances of each training instance it encounters (bengio et al. 2003). the layers create intermediate representations of the data that allow the model to “learn” a distributed representation of elements within each instance (e.g., a distributed representation of words within a sentence). this ability of the model to learn requires no (or at most, quite simple) manual feature design. each unit in each layer is connected to each unit in the adjacent layers. vector representations of the data are received by an input layer and transformed in “hidden” layers. the hidden layers feed into a final logistic function layer (i.e., softmax) that outputs a prediction of each possible class as a probability between 0 and 1. the connections between layers are represented by learnable weights; the higher the weight the more influence a unit has on the result. since deep learning is supervised the weights are adjusted with feedback from the gold standard.2 this is done via stochastic gradient descent or some similar optimization algorithm (goldberg 2017) and backpropagation, which tells the model how to change the parameters which build the representation of each layer from the previous layer (lecun, bengio & hinton 2015). computational morphology for language description and documentation 9 figure 5. neural networks, or deep learning, models learn what features in the data are important for giving the correct output until recently, neural networks had the same great disadvantage that unsupervised learning has – superior performance required a great deal of data. data from ldd would have been considered inadequate to train neural networks (duong 2017). even now, a non-neural model can outperform any given neural model that is not tuned to low-resource settings and neural models can be difficult to optimize and tune for low-resource settings (popel & bojar 2018). new methods are being explored to overcome neural models’ dependence on large corpora. examples include fine-tuning a model to the specific task and input data, training intermediate steps, or augmenting the training data. van biljon et al. (2020) looked at fine-tuning a model and determined that shallowor medium-depth size transformer models, for example only 3 encoder and 3 decoder layers, give better results with limited training data. an example of an intermediate training step would be first training a segmentation model to produce surface segments (morphs) and from them to learn underlying forms of morphemes (e.g., “impossible” à “in-possible” à “neg-possible”) (cotterell, vieira & schütze 2016; liu et al. 2018; moeller et al. 2019). the third successful method is augmenting training data. augmentation can be done with artificial word forms (liu et al. 2018) or with information extracted from other resources such as grammars and dictionaries. these are just a few of techniques that have been investigate; there are probably many more that we have not yet discovered. although nlp research in lrl has been growing since the mid-2010’s, very little of it has been applied to linguistic on under-documented languages. one exception is the aggregation project (bender 2014) which has used igt to automatically infer grammatical structure for multiple languages (lepp, zamaraeva & bender 2019; wax 2014). much of their data comes from colorado research in linguistics, volume 25 (2021) 10 the online database of interlinear text (lewis & xia 2010, odin) which is a collection extracted from published linguistic articles or books. these igt excerpts differ from igts produced by field linguists in at least one important way. noise (i.e., typos, inconsistencies, etc.) is generally removed before publication, so that odin does not have the level of noise that field igt does which simplifies pre-processing and does not distract machine learning models with spurious patterns. 4. morphological analysis morphological analysis is a key activity in ldd. morphological analysis is particularly important when working with morphologically complex languages. languages that build words from multiple morphemes or via significant morphophonological changes produce a high number of inflected and compound words which appear to the machine as brand new, unrelated words (dreyer & eisner 2011; goldsmith, lee & xanthos 2017; hammarström & borin 2011; kann, cotterell & schütze 2016; ruokolainen et al. 2013). nlp systems that account for morphology can reduce data sparsity caused by an abundance of individual word forms (mccarthy et al. 2019; vylomova et al. 2020) and help mitigate bias in training data (zmigrod et al. 2019). computational morphological systems have often been limited to languages with publicly available structured data, for example, tables of inflectional patterns in online dictionaries like wiktionary. unfortunately, complete inflectional tables are not easily available for many of the world’s languages. morphological analysis can be separated into two core tasks (cotterell et al. 2015; hammarström & borin 2011; nicolai & kondrak 2017; palmer 2009). the first task is identifying morphemes by determining their shapes and marking boundaries between them, as was done for the lezgi noun in example 1b below. this is known as (unlabeled) morpheme segmentation (creutz & lagus 2007; snyder & barzilay 2008). the second task is deducing each morpheme’s meaning, which is known as parsing, or sometimes called morphological analysis by itself.3 this single step is known in linguistics as glossing, and in computational linguists as labeled morpheme segmentation or, merely, labeling, or tagging. together segmentation and glossing make up a significant part of interlinearization in documentary and descriptive linguistics. these two tasks (step 3 of bird and chiang’s workflow on) are often the most detailed analytical tasks undertaken while still in the field. they are also computational morphology for language description and documentation 11 perhaps the most time-consuming tasks, requiring at least as much, and probably more, time than transcription which can take up to 100 hours for each hour of recorded speech. the linguistic information provided by morpheme segments and glosses lays a vital foundation for subsequent descriptive work. many nlp models have been applied to morpheme segmentation and glossing. automatic morpheme segmentation is commonly traced to the early work of harris (1955) and much segmentation research since then has implemented unsupervised learning which he inspired (goldsmith 2001; creutz & lagus 2002; poon, cherry & toutanova 2009). the preponderance of unsupervised models was probably motivated by the difficulty of finding the high quantity and quality manually segmented data needed to train supervised models. lack of sufficient training data is illustrated by a recent supervised segmentation experiment (ansari et al. 2019) which needed to manually segment a corpus before conducting the experiment. in ldd, segmentation and glossing are typically tackled simultaneously. segmentation finds breaks between morphemes as for the lezgi noun in 1b. glossing labels morphemes with their meaning or function, as in 1c. glossing does not require segmentation, and if done independently is sometimes referred to as parsing. parsing by itself would only provide the information in 1c without indication of morpheme boundaries. (1) a. paçahdin b. paçah-di-n c. king-obl-gen d. ‘king’s’ nlp experiments with lrl often treat segmentation and glossing as separate tasks. other nlp works have taken a tip from ldd and joined the two tasks. joint learning of segmentation and glossing, or labeled segmentation, is less common but has been successful for lrl (cotterell et al., 2015; moeller & hulden, 2018). in general, joint learning is characterized by training on different types of information and is based on the intuition that one type of linguistic knowledge (e.g., syntax) can improve results in another domain (e.g., morphology) (goldsmith et al., 2017). much nlp work focuses on glossing only, which assumes that the data is already is, or does not need to be, segmented into morphemes. mcmillan-major (2020) trained systems to produce a gloss colorado research in linguistics, volume 25 (2021) 12 line by incorporating predictions made from existing segmentations and from the free translation enriched with intent (georgi 2016). both samardzic et al. ( 2015) also used information from other igt lines such as translation and part-of-speech tags to train a system to gloss. 5. inflectional paradigm learning morphology includes the inference of rules that govern word building strategies and the discovery of how word forms are systematically related (roark & sproat 2007). therefore, virpioja et al. (2011) add a third task to morphological analysis: identification of morphologically related words through patterns of inflection. durrett and denero (2013) claim that the inference of inflectional patterns must be based on three assumptions. first, each lexical category is dictated by a subsystem of rules. russian nouns, for example, can be generalized into three simplified patterns of inflection that are usually labeled “masculine”, “feminine”, and “neuter”. lexemes that adhere to the same pattern are grouped into inflectional classes (sometimes called “declensions” for nouns and adjectives and “conjugations” for verbs). the patterns themselves are known as inflectional paradigms. second, inflectional changes are triggered by context and, therefore, the patterns can be inferred from context. descriptive studies look to phonology or else to both phonological structure and the semantic content of the lexeme for the triggering context. computational models, due to the nature of their input, look to orthographic context. the third assumption is that each stem morpheme is inflected consistently according to the inflectional class it belongs to with any idiosyncrasies of the stem itself. monson et al. (2007) give two guiding principles for computational paradigm learning. one is that inflected forms of a lemma look similar to each other. this principle holds well enough to serve as a solid working assumption, although languages abound with exceptions and inflection can even be suppletive (e.g., go vs. went, etc.). the second principle is that “in any given corpus, a particular lexeme will likely not occur in all possible inflected forms”. inflectional paradigms can be quite large. languages may have hundreds or even thousands of forms per lemma (corbett 2013). even with a large corpus, attempts to learn paradigms, like the one illustrated in figure 6, by only using the corpus will leave empty cells in the paradigm’s table. it is possible that certain forms may never occur in natural language even though they are grammatically possible (silfverberg & hulden 2018). additionally, frequent words often follow irregular patterns, as does computational morphology for language description and documentation 13 the english verb be. for these reasons, the ldd workflow includes elicitation of morphological paradigms (lupke 2010; boerger et al. 2016). figure 6. inflectional paradigm of the english verb “to be” computational models can learn frequent and regular paradigmatic patterns with over 90% accuracy even in low-resource settings (hammarström & borin 2011; durrett & denero 2013; ahlberg, forsberg & hulden 2014). most early work on paradigm induction applied unsupervised learning to concatenative morphology (goldsmith 2001; chan 2006; monson et al. 2007). the unsupervised version of the paradigm completion task (jin et al. 2020) has been the subject of a recent shared task (kann et al. 2020), with the conclusion that it is extremely challenging for current state-of-the-art systems. semi-supervised models have been more recently applied on concatenative and non-concatenative languages (dreyer & eisner 2011; durrett & denero 2013). supervised learning has also been applied to inflectional morphology. some work focuses on generating inflected forms, including work motivated by the paradigm cell filling problem (pcfp), illustrated in figure 7 (ackerman, blevins & malouf 2009). the pcfp is framed as an attempt to model how new speakers (e.g., young children ) infer the inflected forms they have not yet encountered (dreyer & eisner 2011; ahlberg, forsberg & hulden 2015; malouf 2016; silfverberg & hulden 2018). colorado research in linguistics, volume 25 (2021) 14 figure 7. illustration of the paradigm cell filling problem (silfverberg and hulden, 2018) with spanish verb paradigms other work with supervised learning has attempted to induce inflectional paradigms from text. with this method, paradigms are completed by finding overlapping patterns from several incomplete paradigms in text. one method does this by abstracting the longest common subsequence of characters in inflected forms of the same lexeme and then clustering words with same or similar patterns (ahlberg, forsberg & hulden 2014; ahlberg, forsberg & hulden 2015). this is illustrated in figure 8. exceptions or irregularities in the paradigms can be accounted for by collapsing the similar patterns. the experiment has been quite successful for a few indoeuropean languages (german, spanish, catalan, french, galician, italian, portuguese, russian), as well as maltese and finnish. kann et al. (2017a) differed from other approaches in that they encoded multiple inflected forms of a lemma to provide complementary information in order to generate unknown forms. cotterell et al. (2017) introduced neural graphical models which completed paradigms based on principal parts. computational morphology for language description and documentation 15 figure 8. ahlberg et al. (2015): inducing paradigms the longest common subsequences (lcs) rng or swm are extracted (step 1) and represented as x1 and x2 which replace the lcs (step 2). words with the same inflectional patterns will be identical (step 3) and can be generalized into paradigms (step 4). the remaining characters i, a, u are assumed to be inflectional affixes. most recent work in paradigm induction has been concerned with generation (as opposed to analysis) of inflected words and has focused on morphological inflection or reinflection (durrett & denero 2013; nicolai, cherry & kondrak 2015; faruqui et al. 2016; kann & schütze 2016; aharoni & goldberg 2017). partially building on these, other research has developed machine learning models which are more suitable for lrl and perform well with limited data (kann, cotterell & schütze 2017b; sharma, katrapati & sharma 2018; makarov & clematide 2018b; wu & cotterell 2019; kann, bowman & cho 2020; wu, cotterell & hulden 2021). 6. conclusion morphology comprises word-building properties in human languages and their accompanying (morpho)syntactic phenomena. historically, nlp and “paper-and-pencil linguists” have taken different and sometimes seemingly incompatible approaches to morphology (karttunen & beesley 2005). despite their out-of-sync approaches, both benefit from morphological analysis (cotterell et al. 2015). morphological analysis includes morpheme segmentation, glossing, and learning inflectional paradigmatic patterns. this paper presented the history of computational models for morphological analysis and looked specifically at their application and success when limited training data is available. the work discussed in this paper demonstrate that computational models colorado research in linguistics, volume 25 (2021) 16 and methods can both successfully perform morphological analysis. this success may be of great benefit for the documentation and description of endangered or under-documented languages. references ackerman, farrell, james p. blevins & robert malouf. 2009. parts and wholes: implicative patterns in inflectional paradigms. in analogy in grammar. oxford: oxford university press. aharoni, roee & yoav goldberg. 2017. morphological inflection generation with hard monotonic attention. proceedings of the 55th annual meeting of the association for computational linguistics (volume 1: long papers) 1. 2004–2015. https://aclanthology.coli.uni-saarland.de/papers/p17-1183/p17-1183 (16 january, 2018). ahlberg, malin, markus forsberg & mans hulden. 2014. semi-supervised learning of morphological paradigms and lexicons. in proceedings of 14th conference of the european chapter of the association for computational linguistics, 569–578. gothenburg, sweden: association for computational linguistics. http://www.aclweb.org/anthology/e/e14/e14-1.pdf#page=569 (1 november, 2016). ahlberg, malin, markus forsberg & mans hulden. 2015. paradigm classification in supervised learning of morphology. in proceedings of the 2015 conference of the north american chapter of the association for computational linguistics: human language technologies, 1024–1029. denver, colorado: association for computational linguistics. http://www.aclweb.org/anthology/n15-1107 (18 january, 2019). anastasopoulos, antonios, david chiang & long duong. 2016. an unsupervised probability model for speech-to-translation alignment of low-resource languages. in proceedings of the 2016 conference on empirical methods in natural language processing, 1255–1263. austin, texas: association for computational linguistics. https://aclweb.org/anthology/d16-1133 (2 july, 2018). ansari, ebrahim, zdeněk žabokrtský, mohammad mahmoudi, hamid haghdoost & jonáš vidra. 2019. supervised morphological segmentation using rich annotated lexicon. in proceedings of the international conference on recent advances in natural language processing (ranlp 2019), 52–61. varna, bulgaria: incoma ltd. https://www.aclweb.org/anthology/r19-1007 (25 january, 2020). computational morphology for language description and documentation 17 auer, eric, albert russel, han sloetjes, peter wittenburg, oliver schreer, s. masnieri, daniel schneider & sebastian tschöpel. 2010. elan as flexible annotation framework for sound and image processing detectors. in nicoletta calzolari, khalid choukri, bente maegaard, joseph mariani, jan odijk, stelios piperidis, mike rosner & daniel tapias (eds.), european language resources association lrec 2010: proceedings of the 7th international language resources and evaluation, 890–893. paris: elra: european language resources association. http://dblp.unitrier.de/db/conf/lrec/lrec2010.html#auerrswsmst10 (30 january, 2014). baldridge, jason & miles osborne. 2008. active learning and logarithmic opinion pools for hpsg parse selection. natural language engineering 14(2). 191–222. baldridge, jason & alexis palmer. 2009. how well does active learning actually work? timebased evaluation of cost-reduction strategies for language documentation. in proceedings of the 2009 conference on empirical methods in natural language processing, 296– 305. singapore. http://www.aclweb.org/anthology/d/d09/d09-1031.pdf. bender, emily m. 2014. language collage: grammatical description with the lingo grammar matrix. in proceedings of the ninth international conference of language resources and evaluation (lrec-2014), 2447–2451. http://www.lrecconf.org/proceedings/lrec2014/pdf/639_paper.pdf. bengio, yoshua, réjean ducharme, pascal vincent & christian jauvin. 2003. a neural probabilistic language model. journal of machine learning research 3. 1137–1155. bergmanis, toms, katharina kann, hinrich schütze & sharon goldwater. 2017. training data augmentation for low-resource morphological inflection. proceedings of the conll sigmorphon 2017 shared task: universal morphological reinflection 31–39. https://aclanthology.coli.uni-saarland.de/papers/k17-2002/k17-2002 (16 january, 2018). biljon, elan van, arnu pretorius & julia kreutzer. 2020. on optimal transformer depth for low-resource language translation. arxiv:2004.04418 [cs]. http://arxiv.org/abs/2004.04418 (14 may, 2020). bird, steven. 2009. natural language processing and linguistic fieldwork. computational linguistics 35(3). 469–474. http://dx.doi.org/10.1162/coli.35.3.469 (16 march, 2015). bird, steven & david chiang. 2012. machine translation for language preservation. in proceedings of coling 2012, 125–134. mumbai. colorado research in linguistics, volume 25 (2021) 18 boerger, brenda h., sarah ruth moeller, will reiman & stephen self. 2016. language and culture documentation manual. leanpub. https://leanpub.com/languageandculturedocumentationmanual (22 may, 2020). bowern, claire. 2008. linguistic fieldwork: a practical guide. houndmills, basingstoke, hampshire [england]; new york: palgrave macmillan. chan, erwin. 2006. learning probabilistic paradigms for morphology in a latent class model. in proceedings of the eighth meeting of the acl special interest group on computational phonology and morphology (sigphon ’06), 69–78. stroudsburg, pa, usa: association for computational linguistics. http://dl.acm.org/citation.cfm?id=1622165.1622174. corbett, greville g. 2013. the unique challenge of the archi paradigm. in chundra cathcart, shinae kang & clare s. sandy (eds.), proceedings of the 37th annual meeting of the berkeley linguistics society:special session on languages of the caucasus, 52–67. berkeley, ca. https://citeseerx.ist.psu.edu/viewdoc/download?doi=10.1.1.657.9753&rep=rep1&type=pd f (12 january, 2021). cotterell, ryan & georg heigold. 2017. cross-lingual character-level neural morphological tagging. proceedings of the 2017 conference on empirical methods in natural language processing 748–759. https://aclanthology.coli.uni-saarland.de/papers/d171078/d17-1078 (16 january, 2018). cotterell, ryan, christo kirov, john sylak-glassman, géraldine walther, ekaterina vylomova, arya d. mccarthy, katharina kann, et al. 2018. the conll–sigmorphon 2018 shared task: universal morphological reinflection. in proceedings of the conll sigmorphon 2018 shared task: universal morphological reinflection, 1–27. brussels: association for computational linguistics. http://www.aclweb.org/anthology/k18-3001 (2 november, 2018). cotterell, ryan, christo kirov, john sylak-glassman, géraldine walther, ekaterina vylomova, patrick xia, manaal faruqui, et al. 2017. conll-sigmorphon 2017 shared task: universal morphological reinflection in 52 languages. in proceedings of the conll sigmorphon 2017 shared task: universal morphological reinflection, 1–30. computational morphology for language description and documentation 19 vancouver: association for computational linguistics. http://www.aclweb.org/anthology/k17-2001. cotterell, ryan, christo kirov, john sylak-glassman, david yarowsky, jason eisner & mans hulden. 2016. the sigmorphon 2016 shared task—morphological reinflection. in proceedings of the 14th sigmorphon workshop on computational research in phonetics, phonology, and morphology, 10–22. cotterell, ryan, thomas müller, alexander m. fraser & hinrich schütze. 2015. labeled morphological segmentation with semi-markov models. in proceedings of the nineteenth conference on computational natural language learning, 164–174. beijing, china: association for computational linguistics. cotterell, ryan, john sylak-glassman & christo kirov. 2017. neural graphical models over strings for principal parts morphological paradigm completion. proceedings of the 15th conference of the european chapter of the association for computational linguistics: volume 2, short papers 2. 759–765. https://aclanthology.coli.uni-saarland.de/papers/e172120/e17-2120 (16 january, 2018). cotterell, ryan, tim vieira & hinrich schütze. 2016. a joint model of orthography and morphological segmentation. proceedings of the 2016 conference of the north american chapter of the association for computational linguistics: human language technologies 664–669. https://aclanthology.coli.uni-saarland.de/papers/n16-1080/n161080 (16 january, 2018). cox, christopher, gilles bouliame & jahangir alam. 2019. taking aim at the transcription bottleneck: integrating speech technology into language documentation and conservation. slideshow presented at the 6th international conference on language documentation and conservation (icdlc), honolulu, hi. https://scholarspace.manoa.hawaii.edu/handle/10125/44841. creutz, mathias & krista lagus. 2002. unsupervised discovery of morphemes. in proceedings of the acl-02 workshop on morphological and phonological learning-volume 6, 21–30. philadelphia, pa: association for computational linguistics. creutz, mathias & krista lagus. 2007. unsupervised models for morpheme segmentation and morphology learning. acm trans. speech lang. process. 4(1). 3:1-3:34. http://users.ics.aalto.fi/krista/papers/creutz07acmtslp.pdf. colorado research in linguistics, volume 25 (2021) 20 czaykowska-higgins, ewa. 2009. research models, community engagement, and linguistic fieldwork: reflections on working within canadian indigenous communities. language documentation & conservation 3(1). 15–50. http://scholarspace.manoa.hawaii.edu/bitstream/handle/10125/4423/czaykowskahiggins.p df?sequence=1. dreyer, markus & jason eisner. 2011. discovering morphological paradigms from plain text using a dirichlet process mixture model. in proceedings of the conference on empirical methods in natural language processing, 616–627. association for computational linguistics. duong, long. 2017. natural language processing for resource-poor languages. melbourne, australia: university of melbourne phd thesis. http://minervaaccess.unimelb.edu.au/handle/11343/192938 (29 june, 2018). duong, long, antonios anastasopoulos, david chiang, steven bird & trevor cohn. 2016. an attentional model for speech translation without transcription. in proceedings of the 2016 conference of the north american chapter of the association for computational linguistics: human language technologies, 949–959. san diego, california: association for computational linguistics. http://www.aclweb.org/anthology/n16-1109 (6 july, 2018). durrett, greg & john denero. 2013. supervised learning of complete morphological paradigms. in proceedings of the 2013 conference of the north american chapter of the association for computational linguistics: human language technologies, 1185–1195. atlanta, georgia: association for computational linguistics. https://research.google.com/pubs/pub41850.html (2 january, 2018). faruqui, manaal, yulia tsvetkov, graham neubig & chris dyer. 2016. morphological inflection generation using character sequence to sequence learning. in proceedings of the 2016 conference of the north american chapter of the association for computational linguistics: human language technologies, 634–643. san diego, california: association for computational linguistics. https://www.aclweb.org/anthology/n16-1077 (28 april, 2021). felt, paul. 2012. improving the effectiveness of machine-assisted annotation. brigham young university ma thesis. https://scholarsarchive.byu.edu/etd/3214. computational morphology for language description and documentation 21 foley, ben, josh arnold, rolando coto-solano, gautier durantin, t. mark ellison, daan van esch, scott heath, et al. 2018. building speech recognition systems for language documentation: the coedl endangered language pipeline and inference system. in proceedings of the 6th international workshop on spoken language technologies for under-resourced languages (sltu 2018). https://www.iscaspeech.org/archive/sltu_2018/pdfs/ben.pdf (7 march, 2019). forsberg, markus & mans hulden. 2016. learning transducer models for morphological analysis from example inflections. in proceedings of the sigfsm workshop on statistical nlp and weighted automata, 42–50. berlin, germany: association for computational linguistics. http://anthology.aclweb.org/w16-2405 (18 january, 2019). georgi, ryan alden. 2016. from aari to zulu: massively multilingual creation of language tools using interlinear glossed text. phd thesis. https://digital.lib.washington.edu:443/researchworks/handle/1773/37168 (28 september, 2020). goldberg, yoav. 2017. neural network methods for natural language processing (synthesis lectures on human language technologies 37). morgan & claypool. http://www.morganclaypool.com/doi/10.2200/s00762ed1v01y201703hlt037 (12 january, 2019). goldsmith, john. 2001. unsupervised learning of the morphology of a natural language. computational linguistics 27(2). 153–198. goldsmith, john, jackson lee & aris xanthos. 2017. computational learning of morphology. annual review of linguistics 3. 85–106. hammarström, harald & lars borin. 2011. unsupervised learning of morphology. computational linguistics 37(2). 309–350. http://www.mitpressjournals.org/doi/abs/10.1162/coli_a_00050 (29 august, 2017). harris, zellig. 1955. from phoneme to morpheme. language 31(2). 190–222. he, luheng, julian michael, mike lewis & luke zettlemoyer. 2016. human-in-the-loop parsing. in proceedings of the 2016 conference on empirical methods in natural language processing, 2337–2342. austin, texas: association for computational linguistic. colorado research in linguistics, volume 25 (2021) 22 himmelmann, nikolaus p. 1998. documentary and descriptive linguistics. linguistics (36). 161–195. jin, huiming, liwei cai, yihui peng, chen xia, arya mccarthy & katharina kann. 2020. unsupervised morphological paradigm completion. in proceedings of the 58th annual meeting of the association for computational linguistics, 6696–6707. online: association for computational linguistics. https://doi.org/10.18653/v1/2020.aclmain.598. https://www.aclweb.org/anthology/2020.acl-main.598 (8 april, 2021). kann, katharina, samuel r. bowman & kyunghyun cho. 2020. learning to learn morphological inflection for resource-poor languages. proceedings of the aaai conference on artificial intelligence 34(05). 8058–8065. https://ojs.aaai.org/index.php/aaai/article/view/6316 (26 january, 2021). kann, katharina, ryan cotterell & hinrich schütze. 2016. neural morphological analysis: encoding-decoding canonical segments. in proceedings of the 2016 conference on empirical methods in natural language processing, 961–967. austin, texas: association for computational linguistics. http://aclweb.org/anthology/d16-1097 (6 november, 2020). kann, katharina, ryan cotterell & hinrich schütze. 2017a. neural multi-source morphological reinflection. proceedings of the 15th conference of the european chapter of the association for computational linguistics: volume 1, long papers 514–524. kann, katharina, ryan cotterell & hinrich schütze. 2017b. one-shot neural cross-lingual transfer for paradigm completion. in proceedings of the 55th annual meeting of the association for computational linguistics (volume 1: long papers), 1993–2003. vancouver, canada: association for computational linguistics. https://www.aclweb.org/anthology/p17-1182 (26 january, 2021). kann, katharina, arya d. mccarthy, garrett nicolai & mans hulden. 2020. the sigmorphon 2020 shared task on unsupervised morphological paradigm completion. in proceedings of the 17th sigmorphon workshop on computational research in phonetics, phonology, and morphology, 51–62. online: association for computational linguistics. https://www.aclweb.org/anthology/2020.sigmorphon-1.3 (8 april, 2021). computational morphology for language description and documentation 23 kann, katharina & hinrich schütze. 2016. single-model encoder-decoder with explicit morphological representation for reinflection. in proceedings of the 54th annual meeting of the association for computational linguistics (volume 2: short papers), 555–560. berlin, germany: association for computational linguistics. http://anthology.aclweb.org/p16-2090 (5 november, 2018). karttunen, lauri & kenneth r. beesley. 2005. twenty-five years of finite-state morphology. in inquiries into words, a festschrift for kimmo koskenniemi on his 60th birthday, 71–83. csli publications. kirschenbaum, amit, peter wittenburg & gerhard heyer. 2012. unsupervised morphological analysis of small corpora: first experiments with kilivila. language documentation & conservation special publication (potentials of language documentation: methods, analyses, and utilization) 3. 25–31. http://scholarspace.manoa.hawaii.edu/bitstream/handle/10125/4513/04kirschenbaumetal. pdf. kohonen, oskar, sami virpioja & krista lagus. 2010. semi-supervised learning of concatenative morphology. in proceedings of the 11th meeting of the acl special interest group on computational morphology and phonology, 78–86. association for computational linguistics. lecun, yann, yoshua bengio & geoffrey hinton. 2015. deep learning. nature 521(7553). 436– 444. https://doi.org/10.1038/nature14539. https://doi.org/10.1038/nature14539. lepp, haley, olga zamaraeva & emily m. bender. 2019. visualizing inferred morphotactic systems. in proceedings of the 2019 conference of the north american chapter of the association for computational linguistics (demonstrations), 127–131. minneapolis, minnesota: association for computational linguistics. https://www.aclweb.org/anthology/n19-4022 (6 april, 2020). lewis, william d. & fei xia. 2010. developing odin: a multilingual repository of annotated language data for hundreds of the world’s languages. literary and linguistic computing 25(3). 303–319. http://llc.oxfordjournals.org/content/25/3/303 (29 january, 2014). liu, ling, ilamvazhuthy subbiah, adam wiemerslage, jonathan lilley & sarah moeller. 2018. morphological reinflection in context: cu boulder’s submission to conllcolorado research in linguistics, volume 25 (2021) 24 sigmorphon 2018 shared task. in proceedings of the conll sigmorphon 2018 shared task: universal morphological reinflection, 86–92. brussels: association for computational linguistics. http://www.aclweb.org/anthology/k18-3010 (2 november, 2018). lupke, friederike. 2010. data collection methods for field-based language documentation. language documentation and description 7. 55–104. makarov, peter & simon clematide. 2018a. uzh at conll–sigmorphon 2018 shared task on universal morphological reinflection. in proceedings of the conll–sigmorphon 2018 shared task: universal morphological reinflection, 69–75. brussels: association for computational linguistics. https://www.aclweb.org/anthology/k18-3008 (16 may, 2019). makarov, peter & simon clematide. 2018b. imitation learning for neural morphological string transduction. in proceedings of the 2018 conference on empirical methods in natural language processing, 2877–2882. brussels, belgium: association for computational linguistics. https://www.aclweb.org/anthology/d18-1314 (28 april, 2021). makarov, peter, tatiana ruzsics & simon clematide. 2017. align and copy: uzh at sigmorphon 2017 shared task for morphological reinflection. in proceedings of the conll sigmorphon 2017 shared task: universal morphological reinflection, 49– 57. vancouver: association for computational linguistics. http://www.aclweb.org/anthology/k17-2004 (3 november, 2018). malouf, robert. 2016. generating morphological paradigms with a recurrent neural network. san diego linguistic papers 6. 122–129. mccarthy, arya d, ekaterina vylomova, shijie wu, chaitanya malaviya, lawrence wolfsonkin, garrett nicolai, christo kirov, et al. 2019. the sigmorphon 2019 shared task: crosslinguality and context in morphology. in proceedings of the 16th sigmorphon workshop on computational research in phonetics, phonology, and morphology. florence, italy: association for computational linguistics. mcmillan-major, angelina. 2020. automating gloss generation in interlinear glossed text. in proceedings of the society for computation in linguistics, vol. 3, 338–349. https://doi.org/10.7275/tsmk-sa32. https://scholarworks.umass.edu/scil/vol3/iss1/33. computational morphology for language description and documentation 25 moeller, sarah & mans hulden. 2018. automatic glossing in a low-resource setting for language documentation. in proceedings of the workshop on computational modeling of polysynthetic languages, 84–93. santa fe, new mexico, usa: association for computational linguistics. http://www.aclweb.org/anthology/w18-4809 (22 august, 2018). moeller, sarah, ghazaleh kazeminejad, andrew cowell & mans hulden. 2018. a neural morphological analyzer for arapaho verbs learned from a finite state transducer. in proceedings of the workshop on computational modeling of polysynthetic languages, 12–20. santa fe, new mexico, usa: association for computational linguistics. http://www.aclweb.org/anthology/w18-4802 (22 august, 2018). moeller, sarah, ghazaleh kazeminejad, andrew cowell & mans hulden. 2019. improving lowresource morphological learning with intermediate forms from finite state transducers. in proceedings of the workshop on computational methods for endangered languages, vol. 1. honolulu, hi. https://www.aclweb.org/anthology/w196011/. monson, christian, jaime carbonell, alon lavie & lori levin. 2007. paramor: finding paradigms across morphology. in advances in multilingual and multimodal information retrieval (lecture notes in computer science), 900–907. springer, berlin, heidelberg. https://link.springer.com/chapter/10.1007/978-3-540-85760-0_115 (27 february, 2018). moon, taesun, katrin erk & jason baldridge. 2009. unsupervised morphological segmentation and clustering with document boundaries. in proceedings of the 2009 conference on empirical methods in natural language processing: volume 2, 668–677. association for computational linguistics. nicolai, garrett, colin cherry & grzegorz kondrak. 2015. inflection generation as discriminative string transduction. in proceedings of the 2015 conference of the north american chapter of the association for computational linguistics: human language technologies, 922–931. denver, colorado: association for computational linguistics. https://doi.org/10.3115/v1/n15-1093. https://www.aclweb.org/anthology/n15-1093 (26 january, 2021). nicolai, garrett, kyle gorman & ryan cotterell (eds.). 2020. proceedings of the 17th sigmorphon workshop on computational research in phonetics, phonology, and colorado research in linguistics, volume 25 (2021) 26 morphology. online: association for computational linguistics. https://www.aclweb.org/anthology/2020.sigmorphon-1.0 (8 april, 2021). nicolai, garrett & grzegorz kondrak. 2017. morphological analysis without expert annotation. proceedings of the 15th conference of the european chapter of the association for computational linguistics: volume 2, short papers 2. 211–216. https://aclanthology.coli.uni-saarland.de/papers/e17-2034/e17-2034 (16 january, 2018). palmer, alexis mary. 2009. semi-automated annotation and active learning for language documentation. university of texas at austin phd thesis. palmer, alexis, taesun moon, jason baldridge, katrin erk, eric campbell & telma can. 2010. computational strategies for reducing annotation effort in language documentation. linguistic issues in language technology 3(4). 1–42. http://journals.linguisticsociety.org/elanguage/lilt/article/view/663.html (24 january, 2014). poon, hoifung, colin cherry & kristina toutanova. 2009. unsupervised morphological segmentation with log-linear models. in proceedings of human language technologies: the 2009 annual conference of the north american chapter of the association for computational linguistics, 209–217. association for computational linguistics. popel, martin & ondřej bojar. 2018. training tips for the transformer model. the prague bulletin of mathematical linguistics 110(1). 43–70. http://content.sciendo.com/view/journals/pralin/110/1/article-p43.xml (13 may, 2020). rice, sally & dorothy thunder. 2017. community-based corpus-building: three case studies. presented at the 5th international conference on language documentation and conservation (icldc), honolulu, hi. http://scholarspace.manoa.hawaii.edu/handle/10125/42052 (6 june, 2017). roark, brian & richard william sproat. 2007. computational approaches to morphology and syntax. oxford; new york: oxford university press. rogers, chris. 2010. review of fieldworks language explorer (flex) 3.0. language documentation & conservation 4. 78–84. http://scholarspace.manoa.hawaii.edu/handle/10125/4471 (13 march, 2018). ruokolainen, teemu, oskar kohonen, kairit sirts, stig-arne grönroos, mikko kurimo & sami virpioja. 2016. a comparative study of minimally supervised morphological computational morphology for language description and documentation 27 segmentation. computational linguistics 42(1). 91–120. http://www.mitpressjournals.org/doi/10.1162/coli_a_00243 (17 january, 2018). ruokolainen, teemu, oskar kohonen, sami virpioja & mikko kurimo. 2013. supervised morphological segmentation in a low-resource learning setting using conditional random fields. in conll, 29–37. samardzic, tanja, robert schikowski & sabine stoll. 2015. automatic interlinear glossing as two-level sequence classification. in proceedings of the 9th sighum workshop on language technology for cultural herita ge, social sciences, and humanities (latech), 68–72. beijing, china: association for computational linguistics. http://aclweb.org/anthology/w15-3710 (25 may, 2020). seifart, frank, nicholas evans, harald hammarström & stephen c. levinson. 2018. language documentation twenty-five years on. language 94(4). e324–e345. https://muse.jhu.edu/article/712110 (8 october, 2019). sharma, abhishek, ganesh katrapati & dipti misra sharma. 2018. iit(bhu)–iiith at conll– sigmorphon 2018 shared task on universal morphological reinflection. in proceedings of the conll–sigmorphon 2018 shared task: universal morphological reinflection, 105–111. brussels: association for computational linguistics. https://www.aclweb.org/anthology/k18-3013 (26 january, 2021). silfverberg, miikka & mans hulden. 2018. an encoder-decoder approach to the paradigm cell filling problem. in proceedings of the 2018 conference on empirical methods in natural language processing, 2883–2889. brussels, belgium: association for computational linguistics. https://www.aclweb.org/anthology/d18-1315 (25 april, 2019). snyder, benjamin & regina barzilay. 2008. unsupervised multilingual learning for morphological segmentation. in acl, 737–745. soricut, radu & franz och. 2015. unsupervised morphology induction using word embeddings. in proceedings of the 2015 conference of the north american chapter of the association for computational linguistics: human language technologies, 1627– 1637. denver, colorado. sudhakar, akhilesh & anil kumar singh. 2017. experiments on morphological reinflection: conll-2017 shared task. proceedings of the conll sigmorphon 2017 shared colorado research in linguistics, volume 25 (2021) 28 task: universal morphological reinflection 71–78. https://doi.org/10.18653/v1/k172007. https://aclanthology.coli.uni-saarland.de/papers/k17-2007/k17-2007 (16 january, 2018). vallejos, rosa. 2014. integrating language documentation, language preservation, and linguistic research: working with the kokamas from the amazon. language documentation & conservation 8. 38–65. http://scholarspace.manoa.hawaii.edu/bitstream/handle/10125/4618/vallejos.pdf?sequenc e=1. virpioja, sami, ville turunen, sebastian spiegler, oskar kohonen & mikko kurimo. 2011. empirical comparison of evaluation methods for unsupervised learning of morphology. trait. autom. des langues 52(2). 45–90. vylomova, ekaterina, jennifer white, elizabeth salesky, sabrina j. mielke, shijie wu, edoardo maria ponti, rowan hall maudslay, et al. 2020. sigmorphon 2020 shared task 0: typologically diverse morphological inflection. in proceedings of the 17th sigmorphon workshop on computational research in phonetics, phonology, and morphology, 1–39. online: association for computational linguistics. https://www.aclweb.org/anthology/2020.sigmorphon-1.1 (27 april, 2021). wang, linlin, zhu cao, yu xia & gerard de melo. 2016. morphological segmentation with window lstm neural networks. in aaai’16: proceedings of the thirtieth aaai conference on artificial intelligence, 2842–2848. wax, david allen. 2014. automated grammar engineering for verbal morphology. thesis. https://digital.lib.washington.edu:443/researchworks/handle/1773/25373 (6 april, 2020). woodbury, tony. 2003. defining documentary linguistics. language documentation and description 1. 35–51. wu, shijie & ryan cotterell. 2019. exact hard monotonic attention for character-level transduction. in proceedings of the 57th annual meeting of the association for computational linguistics, 1530–1537. florence, italy: association for computational linguistics. https://www.aclweb.org/anthology/p19-1148 (26 january, 2021). wu, shijie, ryan cotterell & mans hulden. 2021. applying the transformer to character-level transduction. in proceedings of the 16th conference of the european chapter of the computational morphology for language description and documentation 29 association for computational linguistics: main volume, 1901--1907. association for computational linguistics. https://aclanthology.org/2021.eacl-main.163. xia, fei, william d. lewis, michael wayne goodman, glenn slayden, ryan georgi, joshua crowgey & emily m. bender. 2016. enriching a massively multilingual database of interlinear glossed text. language resources and evaluation 50(2). 321–349. https://doi.org/10.1007/s10579-015-9325-4 (14 january, 2020). zmigrod, ran, sabrina j. mielke, hanna wallach & ryan cotterell. 2019. counterfactual data augmentation for mitigating gender stereotypes in languages with rich morphology. in proceedings of the 57th annual meeting of the association for computational linguistics, 1651–1661. florence, italy: association for computational linguistics. https://www.aclweb.org/anthology/p19-1161 (27 april, 2021). colorado research in linguistics, volume 25 (2021) 30 endnotes 1 according to lorelei (https://www.darpa.mil/program/low-resource-languages-for-emergentincidents), ``low-resource'' refers to languages for which no automated human language technology exists. this is due to a lack of linguistic resources. \citet{szymanski_morphological_2012} estimates that 99\% of the world's languages are ``resource-poor''. in linguistics, it is more common to hear other terms. the current work uses the term ``under-described languages'' to refer to languages with minimal published linguistic resources; these have been called ``very scarce-resource language'' \citep{duong_natural_2017}. the term ``under-documented languages'' (duong's ``extremely scarce-resource languages'') refer to languages that lack sufficient raw or annotated data to write a full reference grammar. the term ``endangered languages'' refers to languages that are predicted to have no native speakers within a generation or two. most endangered languages are under-documented and/or under-described, as well as fitting the definition of low-resource languages. the distinctions between the terms are rarely crucial in the current work. in practice, these terms can be used almost interchangeably. 2 deep learning morphological segmentation has been performed on unsupervised texts with some success (wang et al. 2016). 3nicolai and kondrak (2017) subdivide morphological analysis slightly differently, making a distinction between morphological “analysis” and morphological tagging. they describe morphological analysis as a combination of segmentation and labeling, though they later state that “morphological tagging can be performed as a downstream application of morphological analysis” (p. 211), thereby adhering to the same two distinctions described in this section. microsoft word berlova-cril2021.docx 1 should u rly be txtng ur s/o anya berlova university of colorado boulder this paper examines the impact of phone-based communication and the language of texting on romantic relationships in the united states. texting has become an integral aspect of romantic relationships for many young adults (luo 2014), but expert opinion is divided on the subject. certain studies have shown that texting may lead to “disconnect” and mixed signals, further amplified by the lack of “standardization in [emoji] deployment." conversely, others have found that the similarity of mobile communication between partners may lead to “higher understanding” and greater relationship satisfaction, and emoji usage can be effective in cross-cultural engagement and as a signal of conversational and relational exclusivity. through an analysis of research in media and relationships, the paper argues that the ambiguity presented by texting and phone-based communication can be dangerous, but can also contribute to powerful in-group and pro-social activity. if misused, the effect on romantic relationships can be adverse, so it is crucial that both partners are aware of the possible pitfalls to be able to navigate them with care. keywords: text messaging, new media, technologically mediated communication, cell phone, emoji 1. introduction in the current technological world, the phone has emerged as a popular tool of communication. according to luo (2014:145), “[american] cell owners between the ages of 18 and 24 exchange an average of 109.5 messages on a normal day.” hence, it is logical to assume that texting has become an integral aspect of romantic relationships for many young adults in the united states. however, as discussed by mcmanus (2018), the texting language is viewed as a lesser form of english. furthermore, many linguists and sociologists view texting as a harmful communication tool, with negative effects on both the conversation and the relationship between the parties involved. this paper will aim to answer the question of how the language of texting affects romantic relationships in the united states, and whether it is truly problematic as a language in regard to clarity, understanding, and positive development of the relationship. the first section of the paper details the impact of over-reliance on the texting medium, as well as the positive role texting can serve in relationship formation. the second section of the paper highlights the stylistic features of texting, including emojis and acronyms such as ‘lol’, and their effect on relationship colorado research in linguistics, volume 25 (2021) 2 dynamics. the analysis contributes to literature in linguistics on the social effect of mobile communication within the us. 2. the role of texting in establishing or destabilizing relationships the utilization of texting provides affordances as well as hindrances for romantic relationships. a study by schade et al. (2013) of 276 young adults in the united states examines the lower relationship quality brought about by texting for both men and women, in regard to both texting that is used to work through conflicts or apologize, or texting that is simply too frequent. all of the study participants were either in serious relationships, engaged, or married. study results revealed that texting is a narrow form of expression, so neither side can fully express nor understand the breadth of the emotions of the other person. schade et al. propose that texting may be a safer form of communication, but may serve as a replacement for in-person conversations, leading to disconnect (schade et al. 2013). furthermore, as noted by mcsweeney (2019), the texting component forces people to consider a “new layer of compatibility” when assessing their romantic partner. accurately doing so may be difficult and hinder a relationship’s successful development. mcsweeney (2019) explains that a large amount of information is exchanged over text, but this can open the door to misinterpretation because “people have different language skills, dialects, and even expectations.” hence, developing trust and intimacy may not be easily achieved through the texting platform. the intimacy issue and the dangers of a reliance on the texting medium can further be seen in a study conducted by gershon (2011) regarding media switching and relationships. gershon collected 72 interviews with undergraduates at indiana university, and an issue around media switching had surfaced: more specifically, the lack of transition between texting to in-person conversations. according to one of the interviewees, trill, her relationship fell apart when she was not able to communicate with her romantic interest, todd, in person. the two would text, but when they found each other face to face, “all they did was make out” and never actually talked. trill noted that it was easier for her to text, and at first it seemed to be effective, but eventually, todd became less and less engaged and eventually found someone else (gershon 2011:395). this supports the findings by schade et al. (2013) regarding the emergence of disconnect in a relationship when texting replaces talking in person. should u rly be txtng ur s/o 3 gershon (2011) further points out that texting may be specifically problematic for young adults, who may feel “trapped and frustrated in texting-only relationships.” for instance, another interviewee, rebecca, found herself greatly irritated and upset with her breakup after her boyfriend ended things via text messaging. rebecca wanted “clarity” about his intentions and felt that a text exchange could not fully reflect them. confusion around actions and intentions was also present in the case of another student, halle, whose boyfriend continued to text her after breaking up with her over text. it was clear that for them, a breakup text message meant different things; for halle’s now ‘ex’-boyfriend, texting “did not mean the definitive end of the relationship,” contrary to how it was for halle. in the case of trill, rebecca, and halle, the use of texting as the primary form of communication led to the termination of relationships without the opportunity for reconciliation, which may have been possible if in-person conversation was utilized (gershon 2011). additionally, lefevbre (2017) brings up texting as a route towards the avoidance of confrontation and the usage of ghosting. ghosting refers to “unilaterally ceasing communications (temporarily or permanently) in an effort to withdraw access to individual (s) prompting relationship dissolution (suddenly or gradually) commonly enacted via one or multiple technological medium (s)” (220). as lefevbre (2017) describes, a breakup is achieved more easily through the abrupt disengagement of one partner from an online conversation (such as that over text), but is “negatively endorse[d]” by the recipient of the ghosting. although it is possible that the ghosting initiator does so to “save the non-initiators’ feelings”, the act is viewed as ambiguous and lacking compassion, and a passive approach to a difficult conversation (lefevbre 2017:227). luna (2018) provides an alternative viewpoint into this situation: she concludes that texting can be a tool to bring people closer together, not break their connection. luna (2018) cites trub and barbot (2020) on the motivation behind texting of 982 adults between the ages of 18 to 29. the study revealed that often, people express thoughts over text that they were too shy or anxious to do in person. as stated by trub and barbot (2020), “texting may be used in the service of alleviating fear, anxiety or discomfort related to being in social situations, enabling more confidence and ease in expressing oneself.” a study by reid and reid (2007) offers further support for the usage of texting to build connection. for study participants who carried a greater sense of anxiety, texting was preferred over methods of communication such as voice calls. to them, sending texts felt more comfortable, leading to “expressive and intimate contact” (reid & reid, 2007:433). interestingly, study participants who felt lonely rather than socially anxious preferred colorado research in linguistics, volume 25 (2021) 4 voice calls and viewed texting as a less intimate communication method. perhaps, the asynchronous nature of texting grants a sense of safety, but it may also lead to problems in the way the relationship is viewed. dibble (2017) finds that the edited and perfected nature of a text can lead to “high idealization” of the person on the other side of the screen, causing the communication to seem less tangible (75). the research dibble (2017) describes is related to potential infidelity by partners, and their perception of a side-relationship that takes place online as a “fantasy” rather than reality. even so, it can be posited that a parallel could be drawn to romantic relationships and how they are perceived if the use of texting is heavily utilized. it is possible that the presence of a phone as the moderator between two romantically involved parties can create a schism between real life and the “online” life and distort one’s view of their partner and relationship. more specifically, a relationship may be implicitly viewed in a less serious and more casual way, which opens the door for more ambiguity. furthermore, the aspect of idealization stems from one’s ability to filter a text and portray themselves in the best light, and if this is not displayed in real life, it can create uncomfortable unpredictability. according to murray (1996:1156), one may wish to idealize their partner due to a “desire for security” and wish to “feel safe and secure in one’s commitment.” if there is a lack of consistency in the way a partner is over text versus offline, this can sabotage the feeling of security since the actuality of who the partner is will not be clear (murray 1996). 3. stylistic features of texting the texting medium encompasses unique stylistic and technical means of expression. according to an interview held by turello (2017), emojis are a significant communication tool in the language of texting. the interviewees, which included wendy hall, a computer science professor at the university of southampton, alexandre loktionov, an expert in hieroglyphic texts, and jessica lingel, a social media expert and assistant professor at the university of pennsylvania, come to a consensus that emojis are an efficient means of communicating and can add significant meaning to written words and phrases; however, there is an issue with “standardization in sign deployment” (turello 2017). both loktionov and lingel express concerns about the possibility of misinterpretation due to a lack of concrete “dictionary definitions” for each emoji. if texting is viewed as a separate and individual form of language, then, as with english, “standardization” may not be the accurate approach since various people may have different ways of expressing themselves over text and “texting dialects” may emerge. however, there is a higher danger of should u rly be txtng ur s/o 5 misinterpretation during texting communication since the communicating parties do not see each other in real life, so nonverbal cues or tone of voice cannot be observed and interpreted. even so, as noted by loktionov, the lack of “concreteness” with the use of emojis in texting, in part, has contributed to their growing popularity. according to him, the flexibility of emojis may often add to their usage value due to their ability to convey various emotions and thoughts (turello 2017). furthermore, in the cases where language is a definite barrier, emojis may present an effective method of expressing basic ideas and feelings, especially between people of different language groups. still, loktionov once again points out that this flexibility across various languages makes misinterpretation likely and presents a challenge in clear communication. furthermore, according to lingel and hall, in modern days, since emoji usage is associated most with conveying feelings, they have become “feminized” in american society (turello 2017). this has often led men to reject the use of emojis as a form of linguistic expression, which hinder emojis in becoming a more widely accepted form of language. hence, if a male and a female in a heterosexual relationship are communicating, the female may find herself relying so much more on emojis that it may seem she is speaking a different form of texting dialect. this may serve as a barrier in communication between the two. crystal (2008) provides support for these potential negative aspects by discussing the fact that texting language utilizes a lot of unique features that may lead to confusion and frustration if both parties do not have the same understanding of their meaning. for example, texting uses a lot of omitted letters, which involves the removal of middle or end letters from the word. so, the word ‘message’ may be written as ‘msg’, and the word ‘texting’ may be written as ‘txtin’. furthermore, texting utilizes initialisms, which reduces words to their initial letters, resulting in terms such as ‘jk’ to represent ‘just kidding’ (crystal 2008:42). another feature that is common is the presence of logograms, which translates to “the use of single letters, numerals, and typographic symbols to represent words, parts of words, or even noise associated with actions” (crystal 2008:37). hence, the word ‘for’ may be replaced with the number ‘4’, in a word such as ‘4ever’. the use of newly developed omissions, initialisms, or logograms may not be known by the texting recipient, which may cause dissatisfaction in the communication and, as a result, in the relationship overall. mcculloch (2019) further agrees that communication difficulties and misinterpretations may result if there is not an open conversation about the “means” in which one is expressing one’s thoughts. mcculloch (2019) discusses that texters of different generations may differently colorado research in linguistics, volume 25 (2021) 6 interpret simple features of a message. for example, periods at the end of sentences could be viewed by some as an indication of passive aggression, while others would not give them any meaning beyond adherence to rules of punctuation. mcculloch (2019) believes that there is no “one right way” to use language online and various uses are not wrong, but parties need to be open about their texting style. in fact, similarity in texting style between parties may contribute to greater relationship satisfaction. a study of young adults in romantic relationships by ohadi (2018) shows that a larger similarity between two partners in the use of text messaging, as well as the frequency of “initiating and saying hello via text messaging,” leads to a more satisfying relationship. greater similarity may correspond to higher understanding between two partners in regard to their texting behaviors, which confirms the importance of texting clarity in relationships. a perspective i, myself, have developed is the fact that the features used over texting may be a form of slang; hence, a way to signify ‘in-group status’ (mattiello 2008). thus, it may be perfectly fine that certain logograms or initialisms, for example, would not be understood by everyone who looks at the message. in fact, a couple may develop certain initialisms or omissions themselves that only they understand between each other in order to discuss certain topics efficiently or prevent others from understanding the meaning of their conversation (especially in the cases if these topics would be considered ‘taboo’). in this scenario, the two people in the relationship would form their own in-group, with the texting slang they use signaling their belonging in the relationship and their togetherness. thus, their ability to express their thoughts may not necessarily decrease, as they may have developed certain linguistic replacements for complex ideas that they now share and use among themselves. for instance, gershon (2011) provides an example between two roommates, who developed a certain texting style to indicate different emotions. to convey a friendly tone, they would text each other ‘heyy’ with two y’s, and if only one ‘y’ was used, this signaled negative feelings. the roommates found this an effective means of making their emotions known without needing to say it more explicitly (gershon 2011:399). the usage of texting slang as such may be problematic in the case of the two partners belonging to different in-groups from which they draw their slang. if this is the case, it would be more difficult for them to reach an understanding and would hinder their closeness and connectedness. even so, it is important to note that many features of texting can be viewed as informal, so may make the interaction seem more casual and less meaningful. hence, thoughts that could be should u rly be txtng ur s/o 7 expressed in a deeper or more meaningful way in person would be reduced to a quick and concise format. luo (2014) conducted an online study of 395 participants who described and discussed their texting behavior. the study shows that when partners start relying on texting as their primary form of communication, it may further weaken their attachment and lead to a significant decrease in relationship satisfaction. texting may “reduce the feelings of love, closeness, and connection”, and may amplify miscommunication and misunderstanding (mcmanus 2018). according to mcmanus, acronyms such as ‘lol’ can be used as a signal of passive-aggressiveness or lack of seriousness of the statement to which it corresponds. although this may lighten the mood in certain situations, it may also take away from the power and importance of a text, once again diminishing its meaningfulness. however, it is possible that deep conversation is not necessarily expected to occur over text, since, as mcmanus notes, the language of text emerged to fulfill the need of expressing sufficient-enough emotion using as few letters as possible, with emojis serving as virtual replacements for non-verbal dialogue and tone of voice. mcculloch (2019) agrees that in certain circumstances, emojis can help contextualize the meaning of a text, such as in the cases that sarcasm is intended. mcculloch refers to emojis as “gestures” rather than a form of language, and points out that they are “expressive tools for informal writing” that may serve as an efficient tool to convey attention and irony when doing so with one’s voice is impossible. by using emojis, one can be clear when one is utilizing a playful spirit and “offer deliberate cues to the feelings, emotions, and intentions” behind the text. mcculloch (2019) notes this can be particularly useful in the situation when double meaning is intended. gershon (2011) also discusses that the texting platform can be used to one’s advantage when expressing thoughts or emotions one would have difficulty with face to face. this may include conversations involving anger or jealousy, in which case some people use texting to limit and conceal their emotional intensity. this was also confirmed in a study by pettigrew (2009), who found that individuals can use texting to hide their feelings as well as discuss subjects they would find uncomfortable in-person. or, as indicated by students in gershon’s study, texting can be used as a convenient flirting tool, as a phone potentially alleviates anxieties and grants higher levels of comfort. colorado research in linguistics, volume 25 (2021) 8 4. conclusion in summary, texting as a language form needs to be navigated carefully; there has definitely been evidence of its negative effects on romantic relationships in the united states, but it also has potential to be used as a tool to build connection and efficiently convey information. although schade et al. (2013) provide research to support the idea that texting can lead to misinterpretation, lack of effective communication, and disconnect, there is alternative evidence that shows it is not necessarily so. in the interview by turello (2017), both loktionov and lingel agree that emojis are useful in substituting for real-life emotions so can be good conversational cues. although there is no one “emoji dictionary” and emoji usage is open to interpretation, it can be a good method of conveying thoughts and emotions in situations where language may be a barrier. on the other hand, it is important that emoji users are aware of emojis being “feminized” in our society (turello 2017) and support males in using them in order to prevent major communicational differences between males and females in heterosexual relationships. furthermore, as discussed by crystal (2008), although in certain cases, the usage of texting initialisms, omissions, and logograms may act as another source of confusion and misinterpretation, in my opinion, it is a form of slang that can be used to signify in-group status (mattiello 2008) between the communicating parties and build on their connection. furthermore, although mcmanus (2018) notes that the informal structure of texts can reduce their meaningfulness, it is important to note that texting is not necessarily the tool that is widely used for extremely meaningful communication; its value lies in their ability to quickly and efficiently convey a thought or emotion. overall, it does not seem that there is a concrete answer to whether the effects of texting on romantic relationships in the us are positive or negative. more research in this field from the perspective of socio-linguistics may be useful in coming to this conclusion, particularly with consideration of dialectal or cross-cultural differences. as of right now, i believe that the language of texting may benefit romantic relationships if used with care, but may hinder them if neither side is aware of the negative effects they can have and approaches texting carelessly. references crystal, d. (2008). txtng: the gr8 db8. oxford: oxford university press. gershon, i. (2011). breaking up is hard to do: media switching and media ideologies. journal of linguistic anthropology. should u rly be txtng ur s/o 9 luna, k. (2018, august 9). it’s complicated: our relationship with texting. american psychological association. retrieved from https://www.apa.org/news/press/releases/2018/08/relationship-texting lefebvre, l. (2017). phantom lovers: ghosting as a relationship dissolution strategy in the technological age. in n.m. punyanunt-carter & j.s. wrench (eds). the impact of social media in modern romantic relationships. london: lexington books. luo, s. (2014, april). effects of texting on satisfaction in romantic relationships: the role of attachment. elsevier. mattiello, e. (2008). an introduction to english slang: a description of its morphology, semantics and sociology. polimetrica, international scientific publisher. mcculloch, g. (2019, july). is the internet killing language? lol, no. vox. retrieved from https://www.vox.com/the-highlight/2019/7/22/20702335/internet-language-text-emojisgifs-bad-for-english mcmanus, n. (2018, february). listen up: the linguistics of texting. wellesley centers for women.retrieved from https://www.wcwonline.org/women-s-review-of-bookssept/oct-2018/listen-up-the-linguistics-of-texting mcsweeney, m. (2019, september). revealing your emoticon side: how digital technology has changed the way we talk to each other. cbc radio. retrieved from https://www.cbc.ca/radio/spark/revealing-your-emoticon-side-how-digital-technologyhas-changed-the-way-we-talk-to-each-other-1.5272103 murray, s. (1996). the self-fulfilling nature of positive illusions in romantic relationships: love is not blind, but prescient. journal of personality and social psychology. ohadi, j. (2017, september). i just text to say i love you: partner similarity in texting and relationship satisfaction. elsevier. pettigrew, j. (2009, august). text messaging and connectedness within close interpersonal relationships. marriage and family review. punyanunt-carter, n., & wrench, j. (2017). the impact of social media in modern romantic relationships. london: lexington books. reid, d. & reid, f. (2007, june). text or talk? social anxiety, loneliness, and divergent preferences for cell phone use. cyberpsychology and behavior 10(3). colorado research in linguistics, volume 25 (2021) 10 schade, l., et al. (2013, october). using technology to connect in romantic relationships: effects on attachment, relationship satisfaction, and stability in emerging adults. journal of couple & relationship therapy. trub, l. & barbot, b. (2020) texting – great escape or path to self-expression?: development and validation of the messing motivations questionnaire. measurement and evaluation in counseling and development 53(2). turello, d. (2017, june 15). emoji, texting and social media: how do they impact language? library of congress. retrieved from https://blogs.loc.gov/kluge/2017/06/emoji-textingand-social-media-how-do-they-impact-language/ narrative and identity construction among ethiopian immigrants colorado research in linguistics. june 2007. vol. 20. boulder: university of colorado. © 2007 by weldu michael weldeyesus. narrative and identity construction among ethiopian immigrants weldu michael weldeyesus university of colorado the main objective of this study is to analyze narratives by ethiopian immigrants in the denver metropolitan area, as they share their immigrant experiences while attempting to integrate into the host culture. more specifically, this paper attempts to see how ethiopian immigrants use narrative as a vehicle for constructing their identity as mainstream citizens in the united states. focusing on the issue of language socialization, this study investigates the contrast between a former and a current self exhibited in the narratives and describes the sources of the disparity between these two identity positionings. two main issues are addressed. the first is constructing the current self as a more socialized individual, in contrast with the former-self, representing a less socialized one characterized by linguistic insecurity, nostalgia, and lower self-esteem, among other things. this is exhibited through humorous recall, laughter, code-switching, and at times explicitly stating how one is different currently from who she/he was earlier. the second is the identity that less socialized immigrants construct through negotiation with more socialized immigrants or citizens of the host country making narrative a collaborative enterprise. constructed around linguistic disfluency, these narratives work to project a more assimilated self who is fluent and capable both linguistically culturally. 1 background there are huge numbers of ethiopian immigrants who have come to work and live in the united states. estimates put the number between 300,000-600,000 with a large concentration in the washington dc and maryland area, and the los angeles area in california. there is also a significant ethiopian immigrant population in the denver metro area, which is estimated to be between fifteen and twenty thousand, and the number has been constantly growing. a common problem ethiopian immigrants face, especially at the initial stages of their arrival in the united states, is the acquisition of the english language. even though the ethiopian educational system teaches english from primary up to tertiary (college and university) level and uses it as a medium of instruction from grade seven up to institutions of higher learning, using english as a communication medium for only academic purposes does not offer sufficient exposure to the language or motivation to learn it (dittmar and stutterheim, 1985). when immigrants come to the us, they need at least the minimal degree of proficiency to be able to integrate into the society or to interact with people using english as a medium of communication. stevens (1994) points out that proficiency in english is not only desirable to immigrants for social reasons, it is also necessary for them to access american political and economic life on practical grounds. research in the areas of immigrant studies and intercultural communication indicates that language competence, among other things like cultural distance and level of education, is a major factor that determines the life of immigrants (redmond 1999; nesdale and mak 2003). even if there are numerous immigrants of different national and ethnic origins in the us, not much research has been conducted regarding narrative and identity among 1 weldeyesus: narrative and identity construction among ethiopian immigrants published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) immigrants in general. in connection with this, anna de fina writes, “aside from mainstream images of who immigrants are, little research has been done on the identity that immigrants themselves build and project, and on the processes that affect the formation of such identity” (2000: 132). this study attempts to contribute to filling this gap. 2 objective of the study the main object of this study is to analyze narratives of ethiopian immigrants in denver, colorado, as they tell and retell their immigrant experiences while attempting to integrate into the host culture. more specifically, this paper attempts to see how the ethiopian immigrants use narrative as an optimal vehicle for constructing their identity as mainstream citizens in the us (schiffrin 1996, johnstone 1996, kerby 1991, riessman 1993, inter alia). focusing on the issue of language socialization, this study in particular investigates the identity contrast that is exhibited in the narratives between what i am calling the less-socialized former self and the more-socialized current self and describes the sources of the contrast between these two identity positionings, which are both linguistic and cultural in nature. six narratives of varying lengths have been used as sources of data, two of which were narrated in tigrinya, an ethiopian semitic language that i speak natively, and four of them narrated primarily in english. the narratives fall under the broad genre of ‘immigrant narratives’ especially with regard to the issue of language socialization since they center around the experiences immigrants undergo in attempting to socialize to the overall way of life of the host country both linguistically and culturally. following elinor ochs and bambi scheffelin (1986), i take language socialization to mean both socialization through the use of language and socialization to use language. the expression ‘immigrant narratives’ has been used by de fina (2000) in a nontechnical sense. her focus was on the role of ethnicity in identification of the self among hispanic immigrants. in this paper, the expression ‘immigrant narratives’ will be used to indicate a macro-genre within which the narratives told by ethiopian immigrants would fall. the immigrant narratives i examine here focus on the difficulty that immigrants face in attempting to socialize linguistically and culturally, a shared experience that most immigrants pass through. 3 analyses of narratives in the narratives employed for this study, there is interplay between language, narrative and identity. for this study, i take identity broadly to mean the social positioning of the self against other (bucholtz & hall 2005: 586). the immigrants were using narratives to show how they are different now, i.e., current-self, from who they were in the past, i.e., former-self, particularly when they first came to the united states. the former-self is characterized among other things by linguistic insecurity, feeling of inferiority, asking for help and seeking comfort when faced with challenging circumstances, at times weeping when unable to cope with situations, resisting interacting with people speaking languages other than one’s native language, and a tendency not to initiate conversation. above all, these immigrants reported feeling uncomfortable when 2 2 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/2 doi: https://doi.org/10.25810/f4kj-bz84 narrative and identity construction among ethiopian immigrants confronted with situations that were alien to what they were accustomed to. other features include nostalgia and in extreme cases, a desire to go back to the home country, although this is an option which is rarely taken. the current self, on the other hand, is characterized by much better language proficiency and in some cases native-like fluency, a feeling of equality, self-reliance and self-confidence, dealing with diverse situations and resisting routine, and laughter when talking about past experiences, especially incidents in the first few months of stay in the host country. also, the current self does not limit socializing with people of the same origin but interacts with people of different backgrounds and aspires to be naturalized and become an american citizen. in each of the narratives that i will discuss here, speakers jokingly recall a moment of linguistic mishap or misunderstanding with an american interlocutor. as such, linguistic miscommunication becomes a metaphor for the former unsocialized self, which is presented in contrast with a current and more enlightened socialized self. the speaker in excerpt 1, for example, recalls a previous misunderstanding between herself and a customer when she was working in a donut shop, where she worked in the first two years of her arrival in the us. (1) what size? 1 teki: ninetee:::n, ninety-four it was 1994, january. 2 january. 3 ((everybody laughs.)) 4 month of ja-ha-nuary it was the month of january. 5 ʔɨyya nəyra ihihi ((laughter)) 6 ((all others laugh too.)) ((all others laugh too.)) 7 fevy: ehe okay. 8 teki: dunkin donut yɨsərrɨħ nəyrə, i was working at dunkin donuts, … --> 50 teki: can i have croissant? ʔilunni. he said, “can i have croissant?” --> 51 ihihi “what size.” i said, “what size?” 52 ((reporting her own speech.)) 53 ((everybody laughs.)) --> 54 “what size,” ihi ʔiləyyo. “what size?” i said to him. --> 55 “what size,” ((laughter)) “what size?” ((laughter)) --> 56 “no kɨcroissant croissant.” “no kɨcroissant croissant” 57 ((reporting what the man said.)) 58 əh ok ʔɨhɨm “ok. ehm.” 59 (she nods her head to mean she 60 understood.)) --> 61 yea but what size. “yea, but what size?” 62 fevy: gɨn tay ʔɨyyu but what is croissant? 63 tu-croissant 64 ʔanəwwɨn i don’t know it either. 65 ʔayfələt’kuwwon ʔɨkko 66 zɨnəgərɨzi. 67 teki: zɨbɨllaʕ it is something edible. 68 fevy: ( ) 69 teki: kəriʔəki ʔɨyyə i will show you at safe way 70 safe way. 71 getch: ( ) --> 72 teki nɨssom zɨbɨllaʕ ʔɨyyom what they are asking me is --> 73 zɨbluni zəlləwu. something to eat. 74 fevy: ʔɨwwə. yea. 3 3 weldeyesus: narrative and identity construction among ethiopian immigrants published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 75 teki: ʔanə dɨmma, but i thought that they --> 76 coffee məsilunni? were asking for coffee. the command of english of the narrator of this story was unsteady at that point and she says the training she received was so formulaic that she was told to respond to people who order coffee by saying ‘what size?’ as can be seen in lines 50-61. what is humorous about this excerpt is that the speaker says ‘what size?’ in response to a customer ordering a croissant, a word that she is not familiar with. the miscommunication was eventually resolved by nonverbal means. significantly, this was a very emotional moment for the teller, who says that this is one of the most terrible but memorable experiences she had ever had in the us. she later narrates that after this incident, she went to the basement of the coffee shop and wept bitterly, especially after one of her co-workers mocked what had happened to other employees. she even recalls that she shared the overall incident to her husband who strongly advised her not to be embarrassed or weep when faced with such kinds of incidents. yet, when telling the narrative 11 years later, she narrates her story with laughter, so as to contrast her current socialized self with a former linguistically (and hence culturally) inept self. in short, now she is a well-socialized and enlightened individual with good command of the english language, a good paying job, and she does not appear to have similar problems any more. at times, immigrants recall misunderstandings with citizens of the host country partly due to language and partly due to some broader cultural asymmetries. in excerpt 2, dawit, an ethiopian who came to the us in the mid 1990s, remembers an interaction between himself and an american who works in a post office. (2) get out of here! 1 dawit: you remember once, i was eh:::: applying for employment in 2 post office, i was giving my application, ... 3 and then the lady was eh checking the amamətə mɨhrət 4 {year of salvation} ((this is to mean ad.)) and then 5 you know the year. 6 yosef: yea. 7 dawit: so she was comparing and at some point there was a big gap. 8 yosef: ehem ... 9 dawit: it’s 1988 in my country. she goes, “are you still in the --> 10 80s?” i said, “yes.” and then she said, “get out of here”. 11 hewan: ihihihi 12 dawit: means like don’t be kidding me, right? 13 hewan: yea. 14 dawit: i said oh what did i do. 15 yosef: ihihi --> 16 dawit: i was collecting everything to get out of there. 17 ((everybody laughs including the narrator.)) ... 18 tedi: dropping your items to leave. --> 19 dawit: sh then she go but what did i do? she goes like, “my god --> 20 you guys are funny. you’re still in the 80s? my god.” and 21 she gave me applications again and i wrote bla bla bla bla. --> 22 then i thought ‘get out of here’ means like ‘don’t be 23 kidding me. 24 tedi: hm::: 25 dawit: so from then on 26 tedi: so that’s the kine {pun} --> 27 dawit: i learnt something. 4 4 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/2 doi: https://doi.org/10.25810/f4kj-bz84 narrative and identity construction among ethiopian immigrants the misunderstanding between the narrator of this story and his american interlocutor can be seen from two perspectives. on the part of the american, the root source of the misunderstanding is conceptual, i.e., the calendar difference between the us and ethiopia. ethiopians follow the julian calendar, which is about 7 years and 8 months behind the gregorian calendar. on the part of the ethiopian, the misunderstanding results from the american’s colloquial expression ‘get out of here!’ dawit interprets this statement literally as meaning ‘leave this area!’ as he states in line 16. but while the former self of the narrative was challenged by this, the current self positions it as a learning experience, talking about it as a mere former challenge with no similar setbacks in the present. on the other hand, the use by the american of the second person plural ‘you guys are funny’ (line 20) instead of the singular form while addressing the immigrant shows that the narrator of the story is identified not as an individual but as a member of a whole group, i.e., ethiopians and more specifically ethiopian immigrants. excerpt 3 is also about a misunderstanding that an ethiopian immigrant faced due to linguistic setback, i.e., pronunciation. here dawit recalls a moment where he confused the word ‘boss’ with the word ‘bus’ (lines 7-8): (3) boss versus bus 1 dawit: may be you guys have heard this. i was working eh i was 2 working in san jose parking area again. so one guy came 3 early very early seven a.m. he was working for his boss for 4 his alek’a {boss}. he parked his car. he said, “what’s up 5 my friend.” i said, “good morning.” because i was very 6 ch’əwa {well mannered} at that time. ( ) i was honest i’m --> 7 like, “hei good morning.” he goes, “you know what? my boss --> 8 is late.” i said, ‘what number.’ i thought he said ‘bus.’ 9 you know. 10 tedi: woo::: ahaha ((exaggerated laughter)) 11 dawit: i said what number. ahaha imagine my alək’a {boss} is late. --> 12 sɨnt kut’ɨr? sɨnt kut’ɨr aləka? {what number? boss number what?} 13 ((everybody laughs.)) --> 14 tedi: silly guy. 15 dawit: ehm? yea 16 yosef: that was silly. 17 dawit: that was funny man? 18 yosef: ehm? the narrator then switches into amharic (line12) for emphatic purposes so as to make sure everybody gets the meaning. by joking about this linguistic mishap, dawit expresses that this is something that happened to him some years back, before he became a competent and well-socialized ethiopian with a good command of the english language. (silly has a negative connotation, while funny has a positive one.) in like manner, the teller of the story in excerpt 4, tedi, narrates what he faced in the first couple of weeks of his stay in the us. he caught a bus which took him to an area which he had never been before and hence was not able to identity where he was. (he was supposed to take bus 3-a, but he took bus 3-b, which took him in the opposite direction.) in the excerpt given below, there is a disagreement between the immigrant and an american which makes the negotiation of identity somewhat difficult. 5 5 weldeyesus: narrative and identity construction among ethiopian immigrants published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) (4) you don’t belong to america! 1 tedi: let me finish this. ((he says this fast with the intent of 2 getting the floor.)) finally he said, ok. i said, ok he he 3 got mad. ok “so you don know what city you’re from?” ((the 4 old man is asking and expressing his surprise at the same 5 time.)) 6 ((responding to the man, the storyteller says,)) 7 well, we didn’t leave any we didn’t see any village or any 8 any any land, did we. any forest, did we. because in 9 ethiopia, when we when we when we go to other city, you 10 have to see some kind of forest before you get to another 11 city. and he said, “what?” ihi ((a short laughter as he --> 12 speaks)) you know what? you don’t belong to america.” 13 ((everybody laughs including the storyteller.)) 14 dawit: oh my god! 15 you don’t belong to america. ((repeats what the story 16 teller said out of surprise.)) 17 tedi: you don’t belong to america. ((confirming what the old man --> 18 said)) i wish i met him (this time) how i look. you know --> 19 how what i learned what what i experience i have now. 20 but anyways he he talked to the driver of the the taxi i 21 mean the the bus and he let me go back again and then i had 22 to travel on foot. i think may be ten or thirteen miles 23 home later on. so but anyways yea went on foot, because 24 ((there is overlap here which is hard to detect.)) part of the confusion for this storyteller is that he was trying to assume things as if they were in his country of origin (lines 6-11), where it is not common to see adjacent cities. what results is a problem of acceptance and hence identity challenge on the part of the ethiopian immigrant. when the immigrant tries to justify why he is not able to identify the place where he is supposed to get off the bus, the american not only rejects the justification that the immigrant tries to offer, but also tells him explicitly that he does not belong to the host country altogether (line 12). the immigrant brings up these former challenges as a point of contrast with his current identity. he is a different, competent and well-socialized individual at the time of narrating the story, as shown in lines 18-19 where he explicitly states his wish to show the american how different he is now if he could meet him. in their work on the interrelationship between language and identity, bucholtz and hall (2003, 2005) posit a framework of tactics of intersubjectivity involving three paired components. in extract 4, we observe the adequation and distinction pair of tactics (which roughly mean similarity and difference), as well as what the authors call illegitimation (i.e. delegitimacy). the immigrant attempts to claim that he deserves and is competent enough to be in the us which is adequation in intent, but he is using the tactic of distinction in practice. that is the reason why he could not succeed at that point. on the other hand, the american observes that there is a strong disparity between what he sees and the kind of identity that the ethiopian immigrant claims to be competent enough to belong to the us. he not only emphasizes the distinction (difference) but also delegitimizes the identity the ethiopian immigrant claims. as they are in the process of socializing to the culture and different aspects of life of the host country, immigrants at times do not clearly know what to say and how to act in different contexts. as such, these narratives are often integrated into socialization ‘lessons’. problems are created when immigrants simply try to act as though they were in 6 6 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/2 doi: https://doi.org/10.25810/f4kj-bz84 narrative and identity construction among ethiopian immigrants their native country, a situation which they always recall in their narratives. excerpt 5 is a good illustration of this, where mary, a teenage ethiopian immigrant, recalls her response to a compliment given to her by her tutor about the progress she was exhibiting in her knowledge of english words. (5) why else would i learn? 1 mary: məs’iʔa nəyra. m: she had come. 2 məs’iʔa tay ʔilahɨnni, she came and she said to me, 3 “meron lomɨsɨbba “meron, this time 4 more wordtat fəlit’ki, ”you know more words,” 5 ʔilahɨnni. she said to me. 6 ((reporting what a tutor said to her.)) 7 ʔanə dɨmma? ehe and then i ehm --> 8 “nɨmɨntay dɨyyə “why else would i --> 9 zɨmahar zəlləxu?” ehe learn?” ehe 10 ʔiləyya. ((laughter)) i said to her. 11 ((reporting what she responded to the tutor.)) 12 ((everybody laughs with teki more audibly.)) 13 teki: wəyləkə. taybəlki? ihihi t: amazing! what did you say? 14 mary: ehehe m: laugh --> 15 teki: [thank you zəytɨblɨyya nerki, t: why didn’t you say thank you to her? --> 16 mary: [why am i learning. eheh m: that’s why i am learning. 17 gech: ehehe w: laughs. 18 teki: ehehehe t: laughs. when the tutor gives her a compliment, the teenager responds using a rhetorical question, which literally translates into ‘why else would i learn?’ this basically means ‘i am a student and i am supposed to work hard and know more and more words.’ this response would be quite appropriate in an ethiopian context. two of her listeners, who are more socialized than the teenage immigrant, comment on what the teenager said, arguing that she should have said ‘thank you’ instead. more specifically, she was ‘problematized’ (ochs & taylor, 2001) by teki, her mom, in line 15, who takes the responsibility of assisting her daughter so that she would have a better understanding of the host culture. there is also an overlap between what the teenager says and what her mother suggests to her concerning what she has to say when responding to compliments (lines 15-16). this overlap happens since the teenage immigrant takes a defensive position sticking to what she said as though she were in ethiopia. it is interesting that following the reaction from the interlocutors, the teenager code-switches into english when reporting what she said to her tutor. she perhaps does so either to mean that she responded to her tutor in english, not in tigrinya, or to demonstrate her command of english. in either case, the teenage immigrant is negotiating a different identity, a more socialized one. in this jointly produced narrative, there is a conflict between less socialized and more socialized immigrants, since the latter usually tend to ‘problematize’ the former in attempting to help them socialize to the host culture better. a related issue regarding newly arrived immigrants is going to english as a second language (esl) classes. as part of the esl program or personal effort to have working knowledge of english for communicative purposes, a common strategy 7 7 weldeyesus: narrative and identity construction among ethiopian immigrants published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) employed is memorizing certain formulaic expressions which occur frequently. examples of this are expressions such as ‘you know’ and ‘you too’ as in excerpt 6: (6) you too! 1 yosef: ok. le me tell you what she said. somebody came from eh she was 2 workin. this lady she she was working [and 3 david: [which lady, the girl? 4 yosef: her, thel the girl workin here. she was workin and this guy an 5 american guy came and she he talked to her and he said, “where 6 you from?” she said, “from ethiopia.” “oh! really.” ((quoting 7 the american. he is varying his voice while trying to imitate 8 the american man and the ethiopian girl.)) “yea:::h.” ((quoting 9 her)) “ok. oh. you speak you speak good english.” ((quoting the 10 american)) --> 11 “oh! thank you. you too.” ((quoting the girl.)) 12 ((everybody laughs.)) 13 tedi: wow right. [that’s very funny. 14 david: [that’s very it’s beautiful. 15 elsa: liar. 16 tedi: that’s great. 17 elsa: he is lying. don’t listen to him. -> 18 david: i think i said the same thing. in this excerpt, the expression ‘you too’ has been used inappropriately. having been accustomed to the habit of responding to wishes and appreciation made by their interlocutors, immigrants or speakers of english as a second language may use this expression in occasions where it may not be appropriate as a response of reciprocity. although this expression has been reportedly used by the female speaker, david admits that he had said the same thing (line 18). hence, these narratives often function as collaborative productions of a former self, with speakers and listeners jointly producing a former unsocialized self that they have now moved away from. 4 conclusion this study addresses two main issues in the narratives by ethiopian immigrants. the first is constructing the current self, which is a more socialized and assimilated individual, in contrast with the former-self, which was a less socialized one to the us way of life and was characterized by, among other things, linguistic insecurity, lower self-esteem, and nostalgia. this is done through humorous recall, laughter, codeswitching, and at times explicitly stating how one is currently different from who he/she was earlier. the second main issue is the identity that less socialized immigrants construct through negotiation with more socialized immigrants or citizens of the host country who are observed problematizing less socialized immigrants. thus these narratives are a collaborative enterprise, involving members of the community in a trajectory towards a mainstream american identity. constructed around a linguistic disfluency, these narratives work to project a more assimilated self who is fluent and capable, both linguistically and culturally. in her research on hispanic immigrants, anna de fina (2000) finds that humor is largely absent, especially in argumentative stories of immigrants. contrary to this, we find that humor, in so much as it is used to make a fun of a former self, is a crucial component of the narratives shared by ethiopian immigrants. 8 8 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/2 doi: https://doi.org/10.25810/f4kj-bz84 narrative and identity construction among ethiopian immigrants references bucholtz, mary and kira hall. 2005. identity and interaction: a sociocultural linguistic approach. in discourse studies, 7: 585-614. __________. 2003. language and identity. in alessandro duranti (ed.), a companion to linguistic anthropology. malden, ma: blackwell. cobley, paul. 2001. narrative. london and new york: routledge/taylor & francis. de fina, anna. 2000. orientation in immigrant narratives: the role of ethnicity in the identification of characters. in discourse studies, 2: 2 (pp. 131-157). dittmar, norbert and christiane von stutterheim. 1985. on the discourse of immigrant workers: interethnic communication and communication strategies. in teun a. van kijk (ed.), handbook of discourse analysis, vol.4, discourse analysis in society (pp. 125-149). london: academic press. johnstone, barbara. 1996. the linguistic individual: self-expression in language and linguistics. new york and oxford: oxford university press. labov, william. 1972. language in the inner city. chapter 9: the transformation of experience in narrative syntax. philadelphia: university of pennsylvania press. linde, charlotte. 1993. life stories: the creation of coherence. new york: oxford university press. nesdale, drew and anita s. mak. 2003. ethnic identification, self-esteem and immigrant psychological health. in international journal of intercultural relations, 27: 23-40. ochs, elinor and lisa capps. 2001. living narrative: creating lives in everyday storytelling. cambridge: harvard university press. ochs, elinor and lisa capps. 1996. narrating the self. in annual review of anthropology, 25: 19-43. ochs, elinor and carolyn taylor. 1995. the ‘father knows best’ dynamic in dinnertime narratives. in kira hall and mary bucholtz (eds.) gender articulated. new york & london: routledge. ochs, elinor and bambi b. schieffelin. 1986. introduction. in schieffeling and ochs (eds). language socialization across cultures. new york: cambridge university press. redmond, mark v. (2000). cultural distance as a mediating factor between stress and intercultural communication competence. in international journal of intercultural relations, 24: 151-159. riessman, c. kohler. 1993. narrative analysis. newbury park: sage publications. schiffrin, deborah. 1996. narrative as self-portrait: sociolinguistic construction of identity. in language in society, 25: 2 (pp. 167-203). stevens, gillian. 1994. immigration, emigration, language acquisition, and the english language proficiency of immigrants in the u.s. in barry edmonston and jeffrey s. passel (ed.), immigration and ethnicity: the integration of america’s newest arrivals (pp. 163-185). washington, d.c.: the urban institute press. appendix transcription conventions: italics indicate amharic/tigrinya words in the narratives in english, and english words in the narratives in tigrinya. 9 9 weldeyesus: narrative and identity construction among ethiopian immigrants published by cu scholar, 2007 colorado research in linguistics, volume 20 (2007) 10 bold face indicates louder talk or words uttered with more stress or focus. : a colon indicates lengthening of segments (the more the colons, the longer the segment). (.) dots in parentheses mark pauses (more dots indicate longer the pauses). wia hyphen immediately following a letter indicates an abrupt cutoff in speaking. . a period indicates a falling contour. ? a question mark indicates a rising contour. , a comma indicates a fall-rise. “ ” quotation marks indicate somebody else’s speech stated directly by another. [ a left bracket marks the beginning of an overlap. / / slashes indicated phonetic transcription of some of the words uttered. { } curly brackets give english translation of amharic words uttered by the speakers. ( ) single parentheses enclose words that are not clearly audible (or best guesses). (( )) double parenthesis enclose transcriber comments. 10 colorado research in linguistics, vol. 20 [2007] https://scholar.colorado.edu/cril/vol20/iss1/2 doi: https://doi.org/10.25810/f4kj-bz84 colorado research in linguistics 6-2007 narrative and identity construction among ethiopian immigrants weldu m. weldeyesus recommended citation 1 background 2 objective of the study 3 analyses of narratives 3 and then the lady was eh checking the amamətə mɨhrət 4 {year of salvation} ((this is to mean ad.)) and then 5 you know the year. 6 yosef: yea. 7 dawit: so she was comparing and at some point there was a big gap. 8 yosef: ehem ... 9 dawit: it’s 1988 in my country. she goes, “are you still in the --> 10 80s?” i said, “yes.” and then she said, “get out of here”. 11 hewan: ihihihi 1 dawit: may be you guys have heard this. i was working eh i was 2 working in san jose parking area again. so one guy came 3 early very early seven a.m. he was working for his boss for 4 his alek’a {boss}. he parked his car. he said, “what’s up 5 my friend.” i said, “good morning.” because i was very 6 ch’əwa {well mannered} at that time. ( ) i was honest i’m --> 7 like, “hei good morning.” he goes, “you know what? my boss --> 8 is late.” i said, ‘what number.’ i thought he said ‘bus.’ 9 you know. 10 tedi: woo::: ahaha ((exaggerated laughter)) 13 ((everybody laughs.)) --> 14 tedi: silly guy. 4 conclusion locative constructions in lakhota: evidence for/against “universal conceptual categories” in spatial topology colorado research in linguistics. june 2010. vol. 22. boulder: university of colorado. © 2010 by les sikos. locative constructions in lakhota: evidence for/against “universal conceptual categories” in spatial topology les sikos university of colorado at boulder although languages use a variety of methods to express spatial topological relations, it has generally been assumed that the underlying conceptual categories are universal. however, recent cross-linguistic research has challenged the universal conceptual categories hypothesis on a variety of levels. the goals of this paper are two-fold: first, to analyze and describe the basic locative construction in lakhota, a siouan language. second, since lakhota is often thought to break other typological universals, the lakhota data are evaluated against three versions of the universal conceptual categories hypothesis. the preliminary results seem to indicate that neither the strong view nor its successively weaker versions can account for the lakhota data described here. 1. introduction when we use language to describe spatial topological relations, it may appear as if the language maps directly to fundamental physical distinctions that exist in the world. for example, the english terms “in” and “on” seem to correspond to clear distinctions in spatial relations. we use “in” to describe containment relationships like, “a letter in an envelope” and “an apple in a bowl.” on the other hand, we use “on” for contact relationships like, “a cup on the table” and “the cap on the pen.” however, languages can differ considerably in the ways they partition the same semantic domain. korean, for example, uses the term kkita to describe both “a letter in an envelope” and “a cap on a pen,” but nohta for “a cup on the table” and nehta for “an apple in a bowl” (bowerman & choi 2001). a crucial question that arises from this kind of cross-linguistic comparison is whether or not the underlying conceptual categories for spatial topological relations are universal. 1 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 2. the objective of this study is two fold. the primary goal is to analyze and describe locative constructions in lakhota.1 although little has been written to date on this particular construction, it appears that the language does not have a simple locative and instead describes spatial configurations by using a more complex system. the secondary goal is to determine, in at least a preliminary sense, how lakhota expressions of spatial topology might contribute to recent cross-linguistic research on universal conceptual categories in the spatial topological domain. since lakhota is often thought to break other typological universals (rood & taylor 1996; van valin 2001), the locative construction may provide specific evidence against the universal conceptual categories hypothesis (landau & jackendoff 1993; li & gleitman 2002; cf. levinson & meira 2003). three variations of this hypothesis will be discussed below. 1.1. overview of the lakhota locative construction in general, lakhota appears to describe spatial configurations by combining two different kinds of elements in an adverbial phrase: 1. a small contrastive set of “posture/positional verbs” (e.g. »he ‘exist’; »na)z&i ‘stand’; »ja)ke ‘sit’; »ju)ke ‘lie’)2 2. an elaborate adpostional system (e.g. a»ka)l ‘on top of’; ma»hel ‘inside’; i»sakhib ‘beside’) for example, the following utterance describes the spatial relationship between a cup and table via a combination of the general positional verb »he (‘exist’) and the adpostion a»ka)l (‘on top of)’: (1) wi»jatke ki »waglijutapi (ki) ) el a»ka)l »he cup the table (the) there on top of exists npfigure npground the cup is on the table 1 lakhota (also known as teton sioux) is one of five closely related dialects of the siouan language family, and is spoken on the plains of the northern united states and central canada. lakhota can be further divided into regional or reservation-based subdialects: southwest – pine ridge and rosebud; missouri river area – cheyenne river, lower brule, standing rock (rood & taylor 1996). 2 lakhota utterances are transcribed using the international phonetic alphabet, with the following modifications: s, z, and ts are transcribed as s&, z&, and c&, respectively. 2 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 3 the system is complex in that speakers must select from both categories based on some interaction of multiple variables, including: 1. the spatial relationship between the figure (the object being described) and the ground (the reference point) 3 2. the physical characteristics of the figure 3. the physical characteristics of the ground 4. animacy of the figure and/or ground 5. relative distance of the figure from speaker 6. whether or not the speaker identifies himor herself as being “at” the figure 1.2. overview of the universal conceptual categories hypothesis4 generally speaking, studies of spatial language tend to assume that simple piagetian spatial conceptions are both topological and universal. in other words, concepts like containment, contiguity, and proximity are thought to be represented cognitively by semantic primitives like in, on, and near. furthermore, it is generally assumed that individual languages then directly code these primitive concepts in small, closed classes like adpostitions. if this universal conceptual categories hypothesis is correct, crosslinguistic comparisons of spatial adpositions should provide important evidence linking semantic categories to conceptual categories in a way that is relatively uniform across languages. however, several studies done since the mid-90s have begun to challenge certain aspects of the hypothesis, leading to subsequently weaker and weaker formulations (brown 1994; levinson 1994; bowerman 1996 and 2003; ameka & levinson 2003). in a groundbreaking multi-language study, levinson and meira (2003) compared nine unrelated languages and found that there are significant cross-linguistic differences in how the semantic space is partitioned. although levinson and meira acknowledge that their study is more exploratory than conclusive, they state: 3 throughout the paper, i will use italic capitals to represent the figure and simple capitals for the ground. 4 this overview is drawn largely from levinson and meira (2003). 3 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 4. the differences between the languages turn out to be so significant as to be incompatible with stronger versions of the universal conceptual categories hypothesis. rather, the language-specific spatial adposition meanings seem to emerge as compact subsets of an underlying semantic space, with certain areas being statistical attractors or foci. (2003: 485) clearly, this is not an outright refutation of the universal conceptual categories hypothesis. instead, levinson and meira suggest that spatial conceptions may best be treated as hierarchical divisions of semantic space, similar to recent models used to describe the variation seen in basic color terms across languages (see kay & maffi 1999 for more on this model). levinson and meira convincingly argue that cross-linguistic research in semantic typology would be better served by utilizing a consistent set of stimuli depicting a variety of spatial topological relations (see appendix a and appendix b). since lakhota has been argued to challenge several theories of typological universals (rood & taylor 1996; van valin 2001), the data from lakhota locative constructions may offer evidence for or against levinson and meira’s new hypothesis. therefore, i have adopted much of levinson and meira’s methodology for the current study. as will be shown below, the preliminary results described here indicate that neither the strong view nor its successively weaker versions can account for the lakhota data analyzed in this paper. 2. data analyzed and methods used 2.1. elicitation method in order to elicit data for this project, i followed the general methodology used by levinson and meira (2003). over a period of several months, i showed my consultant, della badwound5, a series of line-drawings from melissa bowerman’s topological relations picture series,6 each depicting a topological 5 i am indebted to della badwound, a native speaker of both lakhota (pine ridge dialect) and english, for providing the data that made this study possible. 6 the drawings are originally from bowerman and pederson (2003). through the assistance of dr. david rood, i was able to receive a complete set of drawings from dr. stephen levinson at the 4 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 5 spatial relation with a designated figure (marked with an arrow) and ground (see appendix a). the set of drawings includes a range of relations that are coded in english by a variety of prepositions like on, in, above, under, and beside, as well as complex prepositions such as inside of, on top of, and on the side of. for each drawing, i asked badwound in english: ‘where is the [figure]?’ she then responded in lakhota. variations were often volunteered by badwound, while others i actively probed for. for example, several of the images represent prototypically western cultural objects which lack lakhota translations (or at least badwound did not know of their translations in lakhota). in such cases, we verbally sketched a parallel scenario using other well-known elements. in addition, i often explored variations of images by replacing either the figure, the ground, and/or the spatial configuration, in an attempt to tease out some of the significant patterns. all of our sessions were recorded in digital audio (mp3 format) and transcribed. 2.2. operational definitions languages not only vary in the kinds of markers they use to code topical relations, but also in the way in which they combine different types of markers into more complex systems. for example, certain languages rely strictly on adpositions (e.g. tiriyó), others also use spatial nouns to varying degrees, with or without locative case markers (e.g. basque, trumai), while some incorporate positional verbs (e.g. dutch, ewe, yélî) (levinson & meira 2003: 492). finding ways to compare these kinds of forms and their combinations across languages can be quite problematic. to further complicate the matter, there does not seem to be much consensus in the literature for characterizing many of these markers. ayano (2001) notes that adpositions have not been clearly defined in part-of-speech research. baker (2003) even goes so far as to say that there is a fundamental disagreement in whether adpositions should be considered functional or lexical categories. therefore, before diving into the details of this paper, i will establish some operational definitions. the lakhota locative construction appears to use a combination of two distinct kinds of markers. for the purposes of this study, i have adopted two of levinson and meira’s working definitions to refer to these two kinds of elements: language and cognition group of the max planck institute for psycholinguistics. although the full set includes 71 drawings, this project only covered 47 scenes. 5 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 6. first, i will use the term adpostion as a combination of both semantic and syntactic criteria such that “a spatial adposition is any expression that heads an adverbial phrase of location in the basic locative construction (answers to where-questions)” (levinson & meira 2003: 486). second, the term locative/positional verb (lpv) will be used to refer to a relatively small set of contrasting verbs of location or position. like many other languages, lakhota makes use of verbs like sit (»ja)ke), stand (»na)z&i), and lie (»ju)ke) to express something about the spatial relation between figure and ground. as can be seen from these examples, lpvs are often derived from posture verbs. both of these definitions will be fleshed out in section 3. 2.3. extensional map finally, elicitation drawings that badwound described using a particular adposition were mapped onto a fixed arrangement of the complete set of drawings used in this study (see section 3.2). this method is helpful in two ways. first, it provides a general idea of how lakhota partitions the conceptual realm of spatial topology. a key assumption here is that a set of drawings referred to by a particular adposition represents the extensional category for that adposition. second, the lakhota mappings can then be compared to the mappings of other languages to see where their boundaries converge or diverge. the fixed arrangement of drawings used in this paper is based on one utilized by levinson and meira (2003).7 however, levinson and meira’s array contains 71 drawings. since this study could not cover all of the scenes, i removed the images that did not appear in the data and left the remaining drawings in their original fixed positions. therefore, a comparison of the lakhota pattern to patterns established for other languages may only give us a rough idea of any cross-linguistic similarities or differences in extensional categories. 7 this method was originally used by bowerman (1996). 6 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 7 3. results of the analysis 3.1. lakota locative construction lakhota expresses spatial configurations with a combination of adpositions and locative/positional verbs (lpvs). based on the utterances that were collected in this study, the basic structure of the locative construction is as follows: (np) (np) (np) (d-adv) (adposition(s)) lpv/pred the parentheses indicate that certain elements can be omitted — only the lpv/pred element is obligatory. the (s) indicates that there is no limit (at least in theory) to the number of adpositions. evidence for this basic ordering will be given throughout section 3.1. some possible variations and exceptions to this order will be discussed in section 3.1.5. the following sections look at each of the elements in detail. 3.1.1. the noun phrase (np) in the locative construction8 due to the very specific way in which the data were elicited, the vast majority of utterances contained two noun phrases. for example, for drawing 1 i asked badwound, where is the cup? she responded with: (1) wi»jatke ki »waglijutapi (ki) ) el a»ka)l »he cup the table (the) there on top of exists npfigure npground the cup is on the table here, both the np representing the figure as well as the np representing the ground are present. on the other hand, several examples show that the np that refers to the ground may be omitted: 8 for an in-depth description of lakhota nouns and noun phrases, see rood and taylor (1996). 7 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 8. (2) »wowapi wa) a»ka)l »he book a on top of exists npfigure a book is up there (on the shelf) (3) »wowapi ki c&ha)»blaska wa) a»ka)l »he book the shelf/board a on top of exists npfigure npground the book is up there on a shelf although both (2) and (3) are grammatically correct, the ground (shelf) is assumed in the former while explicitly stated in the later. the following examples, in contrast, require that the ground be omitted: (4) »ogle ki o»tke coat the hangs the coat is hanging (5) »wowapi »eja i»phaxlog »he paper some pierced through exist some sheets are stuck/pierced there in english, the likely constructions would be, the coat is hanging on the wall and some sheets are stuck on the spike. however, it appears as if the lakhota lpvs o»tke (‘hangs’) and i»phaxlog (‘pierced through’), are incompatible with an explicitly stated ground. i n a later example using o»tke, i attempted to get badwound to express the ground explicitly: (6) »hapi »eja o»tke clothes some hang some clothes are hanging (on a line) (7) »hapi »eja ta)»ka)l o»tke clothes some outside hang some clothes are hanging outside (8) ? »hapi »eja wi)»ka el o»tke ? clothes some rope/line there hang ? some clothes are hanging there on a line first badwound gave a ground-less expression in (6), but when pressed she inserted a location (ta)»ka)l, ‘outside’) in (7). we might call ‘outside’ a pseudoground, but it was not the ground depicted in the drawing. finally, i asked her if it was possible to say (8), a literal translation of the english some clothes are hanging there on a line. her feeling was that it may be “grammatically correct,” 8 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 9 but it “sounded funny” because the “line is a given.” in other words, a native speaker would not describe the scene that way.9 the following set of examples shows how the particular spatial relation of an item being worn by someone can be expressed via one, two, or three nps: (9) pe/i»juskic&a ki »u) headband the (she is) wearing (using) npfigure (she is) wearing a headband (10) »wi»c&i)c&ala ki pe/i»juskic&a wa) na»ta el »u) girl the headband a head there wearing (using) np npfigure npground the girl is wearing a headband on her head (11) »wipiaka ki pa»ƒe el »u) belt the waist/abdomen there worn (used) the belt is worn on the waist the »u)-construction seems to be the preferred way to express the concept of items being worn.10 (9) shows the prototypical form, using only a single np. a better translation might be ‘the headband is worn,’ because both the wearer and the ground are assumed. however, with some coaxing i was able to get badwound to explicitly state the figure, ground, and wearer in (10). finally, (11) shows that both the figure and ground can be used without expressing the wearer. clearly, there is a range of acceptable variation, although within certain constraints, as demonstrated by (4) and (5). rood and taylor state that the only obligatory slot in a lakhota sentence is the verb (1996: 453). however, since there were no instances in this dataset where all the nps were omitted, i cannot say for certain whether or not the locative construction requires at least one np. 9 see section 4.1 for limitations of this study, including the possibility that this elicitation tool is the linguistic equivalent of forcing a square peg into a round hole. 10 see non-spatial predicates in section 3.1.4 for more on why lakhota does not describe certain scenes using the basic locative construction. 9 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 10. 3.1.2. deictic adverbs (d-adv) it appears that the locative construction can use the deictic adverb slot to express the spatial relationship that exists between the speaker and the scene she is describing. for example, when i asked badwound to describe drawing 7, she first said: (12) i)»ktomi wa) tic&e) (el) i»jaje spider a ceiling/roof (there) going/moving a spider is going along on the ceiling the deictic adverb badwound originally used was »el, (‘there’). however, when she repeated the sentence, she omitted the »el. in trying to tease out the meaning of this word, i asked her to imagine that the spider was further and further away from her (i also physically moved the drawing up and away). badwound then answered using different d-advs: (13) i)»ktomi wa) tic&e) hel i»jaje spider a ceiling/roof over there going/moving a spider is going along up there on the ceiling (14) i)»ktomi wa) tic&e) »kakhja i»jaje spider a ceiling/roof to way over there going/moving a spider is going along to (a place) way up there on the ceiling the only element that changes between these utterances is the d-adv. since the spatial relationship between the figure and ground did not change, i assume that it was badwound’s perception of her position in relation to the figure that prompted her to use the different d-advs. although this is a relatively simple example, the implications for the role of introspection in locative constructions may be much more complex.11 four d-advs appeared in the data: 11 rood (2003) outlines a hypothesis wherein the choice of certain adpositions is dependent on whether the speaker imagines the scene to be at the speaker’s own location or someplace away from it, giving rise to complex variations. 10 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 11 lel ‘here’ »el ‘there’ (neutral, default term) »hel ‘over there’ (further away than »el) »kakhja ‘to (a place) way over there’ (telic; further away than both »el and »hel) of these, the most common by far was »el, which appears to be a general default term in addition to simply meaning ‘there.’ this is not particularly surprising if we consider the roots of d-advs, which are formed by adding a demonstrative to an adverb or adposition. lakhota has three demonstrative roots: »le ‘this’ »he ‘that’ (a general, default term) »ka ‘that over there’ (further away than »he) according to rood and tayor, »he is the most semantically neutral of these roots, and is the general term that is used once the location of an np has been identified — either by gesture, by using one of the three demonstratives, or periphrasitcally (1996: 456). therefore, it seems likely that the d-advs work in much the same way as demonstratives. once the spatial relationship between the speaker and the scene has been established, the speaker can default to »el, or even omit it completely. 3.1.3. adpositions lakhota has no prepositions or circumpositions, only postpositions. however, in keeping with the operational definitions outlined in section 2.2., i will continue to call the category by the more general term adposition. although some scholars distinguish between spatial nominals, spatial adverbials, basic adpositions, and derived adpostitions, i have chosen to group them all together for two reasons. first, as mentioned in section 2.2, the crosslinguistic boundaries of these categories are quite fuzzy. levinson and meira make a point of including spatial nominals (e.g. ‘top,’ ‘bottom,’ ‘side’) because even though on top of can be separated from the more complex locative adpositional on the top of, the general spatial relation they both express can tell us something about an underlying concept they may share (2003: 486). the second reason i group all spatial-relation terms under the category of adpostion is specific to lakhota itself. many of its adpositions are derived from verbal stems. an even larger number are derived from an adverb with a related meaning, by simply prefixing an i-. ingham states, “in a sense this type of postposition is infinitely derivable, since potentially any adverb, especially one relating to time or space, can form a postposition by means of the prefix i-” (2003: 41). 11 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 12. furthermore, rood and taylor write, “the line between adverbs and postpositions is sometimes difficult to draw, chiefly because the same words are often used both ways” (1996: 452). in short, for the purposes of this paper, adpostion will be used to describe markers for specific spatial relationships that exist between the figure and ground. this is not to say, however, that subtle distinctions between adpositions are not important for this study. as we shall see, we may be able to establish some taxonomic relationships (at least in a preliminary way) among adpositions (see section 3.2.3). the majority of adpositions that appeared in the data are listed below, organized into broad conceptual categories (e.g. in, on, over): in ma»hel ‘inside,’ ‘within’ on a»ka)l ‘on top of’ a»kaxpa ‘covers’ over i»wa)kab ‘above’ (above “head level,” but not necessarily above a ground) under o»xlathe ‘under’ (contact not allowed) i»oxlathe ‘under,’ ‘right under’ (contact ok) i»hukhul ‘down there’ (below “head level,” not necessarily beneath a ground) near khi»jela ‘near’ i»sakhib ‘beside,’ ‘next to’ around o»homni ‘around’ attached i»phaxlog ‘pierced through’ e»ta) ‘from there’ in an attempt to delineate the boundaries of each term, let’s look at some of the more common and/or interesting examples of how lakhota adpositions describe certain spatial relations. in · ma»hel. the adposition ma»hel is used in two somewhat different ways: (15) tha»spa) wa) »wijatke ma»hel »he apple a cup/bowl inside exists an apple is in the cup (16) »s&u)ka ti ma»hel »xpaje dog house inside lies a dog is lying inside the house 12 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 13 (17) c&i»ska wa) wa»ks&ic&a-pha»xi)te ma»hel »u) spoon a dish towel inside exists a spoon is under the towel examples (15) and (16) show the adposition being used in a way that is quite similar to english. both the apple and the dog are described as being within some container (ground). sentence (17), on the other hand, uses ‘inside’ where english would prefer ‘under’ (lakhota can also use ‘under.’ see next section). perhaps one can think of the lakhota term ma»hel as covering a broader semantic space than the english term in. we will come back to this notion in section 3.2. under · o»xlate, i »oxlate, i »hukul . the following example shows that the same spatial arrangement shown in drawing 24 can be expressed with only two of the three adpositions that can be glossed as under: (18) c&i»ska wa) wa»ks&ic&a-pha»xi)te i»oxlate »u) spoon a dish towel under exists a spoon is under the towel (19) c&i»ska wa) wa»ks&ic&a-pha»xi)te i»hukhul »u) spoon a dish towel down there exists a spoon is under the towel (20) * c&i»ska wa) wa»ks&ic&a-pha»xi)te o»xlate »u) * spoon a dish towel under exists * a spoon is under the towel note that o»xlate cannot be used here. what is particularly interesting is that the two constructions that are most similar in form are the least compatible. i was unable to determine why this was so until i compared these utterances to another set of ‘under’ sentences. badwound used all three adpositional forms in describing drawing 1612: (21) tha»pha wa) »waglijutapi o»xlate »he ball a table under exists a ball is under the table 12 badwound had trouble recalling the lakhota word for ‘chair’ so we substituted ‘table.’ 13 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 14. (22) tha»pha wa) »waglijutapi i»oxlate »he ball a table under exists a ball is under the table (23) tha»pha wa) »waglijutapi i»hukhul »he ball a table down there exists a ball is down there (under) the table according to badwound, there is no difference in meaning between (21) and (22), and both are grammatical. perhaps this can be explained as the effect of the adposition-derivation chain mentioned above. however, there may be another explanation that also accounts for the examples that describe drawing 24. i recalled that badwound made certain hand gestures while describing drawing 16 which gave the impression that i»oxlate was somehow “closer” to her than o»xlate. it did not make sense to me at the time, but later reflection lead me to reinterpret i»oxlate as refering to something that might better be translated as ‘right under the table.’ this explanation would also solve the puzzle of sentence (20). since the towel makes contacts with the spoon, ‘right under the towel’ would make perfect sense. an implication of this solution is that o»xlate cannot be used if there is contact between figure and ground. although i have not been able to test this prediction, the notion of contact will become an important feature in section 3.2.3. according to badwound, o»xlate and i»hukul are not perfect synonyms either. when comparing examples (22) and (23), badwound said that the latter does not necessarily imply that the figure is beneath any kind of ground, while the former does. i interpret badwound’s description as meaning that the utterance in (23) sets up the scene almost as a list: “there’s a ball and a table, and the ball is down there (in relation to the speaker, rather than in relation to the table).” on the other hand, (22) seems to specifically describe the fact that the ball is beneath the table. on · a»ka )l, a»kaxpa. the lakhota adposition for ‘on top of’ is used to describe a figure in contact with a flat horizontal ground. again, the notion of contact is significant (see section 3.2.3). the following example is representative of a great many sentences in the data: (24) »wowapi ki c&ha)»blaska wa) a»ka)l »he book the shelf/board a on top of exists the book is on a shelf a variation, however, can occur by adding a second adposition in series with a»ka)l to express the concept of ‘covering’: 14 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 15 (25) mni»huha wa) »waglijutapi a»ka)l »he linen/cloth a table on top of exists a tablecloth is on the table (26) mni»huha wa) »waglijutapi a»ka)l a»kaxpa »he linen/cloth a table on top of covers exists a tablecloth covers the table although we may be tempted to label the term a»kaxpa in (26) as simply an adverb, there is some evidence in support of grouping it with adpostions. rood and taylor note: the line between adverbs and postpositions is sometimes difficult to draw, chiefly because the same words are often used both ways. english adverbs and prepositions show the same kind of interchangeability. ‘come on out from down in under there!’ has six adverb/prepositions in this kind of ambiguous function. (1996: 452) it seems that a»kaxpa in (26) acts as a serial adposition in the same way as the english down in under there, and carries additional spatial information — namely, that the figure completely covers the top of the ground. we have seen that a»ka)l can be used to express the relationship between a figure and a flat horizontal ground, but it can also describe other types of grounds as well: (27) zi)»tkala wa) »wikha) (el) a»ka)l »ja)ke bird a rope (there) on top of sits a bird is sitting on the line (28) wi»c&has&a wa) ti»-aka)l »naz&i man a roof-on top of stands a man stands on the rooftop (29) wi»c&has&a wa) ti»c&he a»ka)l »naz&i man a roof on top of stands a man stands on top of a roof sentence (27) shows a»ka)l being used with a linear (rather than planar) ground, and (28) and (29) show a flat but angled ground. another notable phenomenon in this set of examples is how adpositions can often combine with nouns to form compounds. the form ti»-aka)l (‘roof-on top of’) in (28) is such a compound. 15 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 16. around · o»homni. at first glance, it may appear that the adposition o»homni is used in the same way that around is used in english: (30) mni»huha wi»jakpa phe»tiz&a)z&a) wa) o»homni i»jakas&kab material shiny lamp/fire a around they tied around they tied a ribbon around the candle (31) »c&hu)kas&ke ti o»homni »he fence house around exists the fence is around the house (32) »wipiaka ki pa»ƒe o»homni »u) belt the waist/abdomen around worn (used) the belt is worn around the waist the three examples above show a wide variation in kinds of ground, from small and large inanimate objects (candle, house) to animates (a woman’s waist). in fact, o»homni was the only adposition that appeared in the »u)-construction (the preferred way to express the concept of items being worn) in this dataset. on the other hand, o»homni is not used in certain situations where english uses around: (33) nu)»psioxli wa) »na)pe el »u) ring a hand/finger there (she is) wearing (using) she is wearing a ring on her finger (34) * nu)»psioxli wa) »na)pe o»homni »u) * ring a hand/finger around (she is) wearing (using) * she is wearing a ring around her finger sentences (33) and (34) show that while one can say ‘she is wearing a ring on her finger’ in lakhota, it is ungrammatical to say ‘she is wearing a ring around her finger.’ what is the rationale behind this categorization? according to badwound, sentences (30), (31), and (32) are all grammatical because the figure “goes around” the ground in each. i interpret this as meaning that each of the figures has two ends, one of which traverses space around the ground to meet the opposite end. even the ‘fence’ in (31) can be thought of in this way (as we can also do in english). conversely, a ring is a solid object that does not have this same property. therefore, corresponding english and lakhota adpositions clearly carve out different areas of spatial conceptualization. attached · i »phaxlog, e»ta ) . there are two lakhota adpostions that carry the concept of attachment. the first is i»phaxlog: 16 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 17 (35) »wowapi »eja i»phaxlog »he paper some pierced through exists some sheets are stuck/pierced there (36) wa»hi)kpe wa) tha»spa) wa) i»phaxlog ja)»ke arrow an apple an bore through sits (and is still there) an arrow bore through an apple and is still there (37) wa»hi)kpe wa) tha»spa) wa) i»phaxlog i»jaje arrow an apple an bore through going (and left a hole) an arrow bore through an apple and left a hole i»phaxlog may best be translated as ‘pierced,’ but can be used in two slightly different senses (as can ‘pierce’ in english). in (35) and (36), the figures were pierced and remain stuck to/on the ground.13 i»phaxlog in this sense can be thought of as having a feature of +attachment. on the other hand, the sense of i»phaxlog in (37) is not one of attachment. instead, the figure pierced the ground, passed on through, and left only a hole. this sense of the adpostition does not share the feature of attachment. another adposition that implies attachment is e»ta) (‘from’). it is particularly interesting because (at least in this dataset) it is only found in association with animate figures that grow from a particular ground: (38) c&ha) wa) pa»ha e»ta) i»c&haƒe tree a hill from grows a tree is growing from the hill (39) tha»spa) wa) c&ha) e»ta) o»tke apple an tree from hanging an apple is hanging from a tree (40) tha»spa) wa) c&ha) e»ta) i»c&aƒe apple an tree from growing an apple is growing from a tree apple and tree in the above examples both ‘grow from’ and are ‘attached at’ a particular place on the ground. 13 the word order difference does not seem to be relevant, perhaps because the instrument and undergoer are obvious from the context. 17 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 18. let’s now turn to the final element in the basic locative construction, the predicate. 3.1.4. lpv/predicate levinson and meira note that many languages can encode topological relations with a contrastive set of locative predicates (2003: 486). lakhota appears to be such a language — it uses various kinds of verbs to express something about the spatial relationship, with or without utilizing the adpositions discussed above. a wide range of predicate types appeared in the data: posture verbs, default verbs of existence, positional verbs, and non-spatial predicates. let’s look at each of these in turn. posture verbs. a subset of the verbs that appeared in the data can be categorized as posture verbs: »ja)ke ‘sit’ »na)z&i ‘stand’ »ju)ke ‘lie’ (used with animate figures only) »xpaje ‘lie’ (used with both animate and inanimate figures) o»tke ‘hang’ although many languages use grammaticalized posture verbs to express something about the axial geometry between the figure and ground, they can be utilized in different ways. for example, germanic languages appear to exhibit a continuum: at one end, dutch and german require posture verbs to express the location of an entity, while at the other end, english rarely utilizes posture verbs (see lemmens 2006). before looking at any specific examples, let’s first identify another type of verb. default verbs of existence. lakhota is similar to english in that the posture verb is not usually a required element and can often be replaced with a more general predicate of existence. lakhota speakers, however, must chose between two general predicates depending on the animacy of the figure: 18 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 19 »he14 ‘exist’ (a general default term, used with inanimate figure) »u) 15 ‘exist’ (a general default term, used with animate figure) some examples of how the posture and animacy of figures interact in the locative construction can be seen in the following set of utterances: (41) »mni »ognake ki »waglijutapi a»ka)l »he water bottle the table on top exists the water bottle is on the table (42) »mni »ognake ki »waglijutapi a»ka)l »na)z&i water bottle the table on top stands the water bottle is standing on the table (43) »mni »ognake ki »waglijutapi a»ka)l »ja)ke water bottle the table on top sits the water bottle is sitting on the table (44) »mni »ognake ki »waglijutapi a»ka)l »xpaje water bottle the table on top lies the water bottle is lying on the table (45) * »mni »ognake ki »waglijutapi a»ka)l »ju)ke * water bottle the table on top lies * the water bottle is lying on the table (46) * »mni »ki »waglijutapi a»ka)l »xpaje * water the table on top lies * the water is lying on the table example (41) shows that »he can be used as the general default term for inanimate objects. sentences (42-44) show some of the various posture verbs that can be used with »mni »ognake (‘water bottle’), depending on what axial geometry it has in relation to the ground (e.g. standing on its base, lying on its side). however, the ungrammaticality of (45) indicates that »ju)ke (‘lies’) cannot be used with inanimate objects. 14 not to be confused with its homonym, the demonstrative root »he (‘that’) 15 not to be confused with its homonym, the lpv »u) (‘wear,’ ‘use’) 19 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 20. example (46) is ungrammatical for a different reason. it appears that lakhota differs from english in that it cannot use a posture verb to describe a situation where water has been spilled on a table. in contrast, compare the english sentence be careful, there’s (pooled) water standing on that table. according to badwound, liquids that are not in some kind of container must be expressed as “running or flowing,” even if it is just pooled on a table. animate objects require a slightly different pattern. none of the following figures allows the use of »he as a general default term: (47) * ha»xa) wa) mni (el) »he * fish a water (there) exists * a fish is (there) in the water (48) * zi)»tkala wa) »wikha (el) a»ka)l »he * bird a rope (there) on top of exists * a bird is standing on the line (49) * ho»ks&ila ki »pheta (el) i»sakhib »he * boy the fire (there) beside exists * the boy is beside a fire (50) * i»gmu wa) o»wi)z&a a»ka)l »he * cat a material/rug on top of exists * a cat is sitting on the material/rug on the other hand, compare the above examples with the following set: (51) zu»zec&a wa) c&ha) el »u) snake a stump/wood there exists there is a snake on the stump (52) zu»zec&a wa) c&ha) el aka)l »ja)ke snake a stump/wood there on top of sits a snake is sitting there on top of the stump (53) zu»zec&a wa) c&ha) el aka)l »ju)ke snake a stump/wood there on top of lies a snake is lying there on top of the stump (54) zu»zec&a wa) c&ha) el aka)l »xpaje snake a stump/wood there on top of lies a snake is lying there on top of the stump (55) * zu»zec&a wa) c&ha) el »he * snake a stump/wood there exists * a snake is there on the stump (56) * zu»zec&a wa) c&ha) el aka)l »he * snake a stump/wood there on top of exists * a snake is there on top of the stump 20 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 21 the direct contrast between (51) and (55-56) indicates that animate figures must take »u) rather than »he as a general default verb. another animacy criteria can be seen by comparing sentences (53) and (54). in contrast with an inanimate object like water bottle, here we can see that animate objects like snake can take either »ju)ke (‘lies’) or »xpaje (‘lies’) and still be grammatical. some lexical items also seem to have specific constraints on which posture verb they can select. the constraint appears to have something to do with the physical dimensions of the figure. for example: (57) i)»ktomi wa) tic&e) (el) »ja)ke spider a ceiling/roof (there) sits (if not moving) a spider is sitting on the ceiling (58) * i)»ktomi wa) tic&e) (el) »na)z)i * spider a ceiling/roof (there) stands (if not moving) * a spider is standing on the ceiling (59) tha»pha wa) »waglijutapi (el) i»hukhul »ja)ke ball a table (there) down there sits a ball is sitting down there (under) the table (60) * tha»pha wa) »waglijutapi (el) i»hukhul »naz&i) * ball a table (there) down there stands * a ball is standing down there (under) the table it appears that both i)»ktomi (‘spider’) and tha»pha (‘ball’) can take »ja)ke (‘sit’) but not »na)z)i (‘stand’). comparing these examples to the data for water bottle on table, snake on stump (drawing 23), and boy beside fire (drawing 38), paints the following picture: sit stand lie water bottle x x x snake x - x boy x x x spider x - - ball x - - one possible conclusion that can be drawn from this is that the height-to-width dimension of a figure combines with the “natural” spatial orientations it tends to take in the real world. it appears that this combination plays a key role in which posture verbs can be selected. figures that are long and thin, but have a multiple natural orientations to ground (i.e. water bottles and boys can often be found upright or on their sides), appear to be able to take any of the posture verbs. snakes, on the other hand, are also long and thin but are rarely found completely upright. spiders and balls seem to form another class of objects that can be 21 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 22. thought of as having approximately equal height and width. furthermore, their orientation does not tend to change much in relation to ground. therefore, this combination may preclude the use of either ‘standing’ or ‘lying.’ positional verbs. some of the predicates that appeared in the data provide important spatial or positional information, but cannot be categorized as posture verbs: i»jaskape ‘sticks (to/on something)’ o»kawiƒe ‘floats (on water)’ ka»xwoke ‘floats (in/on air)’ o»wapi ‘imprinted (on something)’ i»kwqke ‘tie (to something)’, ‘attached (to something)’ i»jakas&kab ‘tie around (something)’ the following comparison shows how scenes that we can describe in english using a single positional verb (floats), require two different verbs in lakhota: (61) tha»spa) wa) »wijatke el o»kawi)ƒe apple a cup/bowl there floats16 an apple is floating inside the cup (62) ma»xpija wa) pa»z&ola (el) i»wa)kab ka»xwoke cloud a pointed little hill (there) above floats a cloud is floating (there) above the pointed little hill non-spatial predicates. levinson and meira note that several languages in their sample (e.g. lao, yukatek) did not use locative constructions when describing certain kinds of scenes. instead, they express these relationships in some other way (either utilizing the resultative or some other descriptive mode) “suggesting that languages perhaps differ in what they consider a fundamentally spatial arrangement” (2003: 495). 16 alternate translations include ‘floating, ‘sailing,’ and/or ‘bobbing up and down (in some liquid)’ 22 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 23 similarly, lakhota appears to prefer non-spatial constructions for several scenes. since these expression evoke a construction other than the basic locative, i will only list the verbs here and indicate which scenes they described: i»jaje ‘goes’ 7, 11, 19, 30 i»c&haƒe ‘grows’ 17, 27, 41 o»nuwe ‘swims’ 32 »u) ‘wears’ (‘uses’) 5, 10, 21, 42, 46 »os&ta ‘put on (clothing)’ 21 i»juthe ‘tried on (clothing)’ 21 c&ha)»nu)pe ‘smokes a cigarette’ 39 »u)pe ‘smokes’ 39 wa»je ‘i did’ 9 as we shall see in the following section, all of the predicates discussed here in section 3.1.4 seem to be able to fill the same slot in the locative construction; therefore i have labeled the slot lpv/pred for locational, postural, and positional verbs (lpvs) , as well as other predicates. 3.1.5. ordering of elements clearly, much more data from multiple speakers will eventually be required to get a more complete picture of the overall patterns that lakhota allows. however, we can make a preliminary summary of the basic locative construction as seen in this data: (np) (np) (np) (d-adv) (adposition(s)) lpv/pred throughout section 3.1 we have explored the multiple patterns that are represented in the formula above. we have looked at each of the elements in detail, as well as identified which are optional and which are obligatory. we have also noted the acceptable combinations of adposition + lpv, acceptable orderings, and identified multiple animacy criteria. a couple of questions still remain, however. for example, the set of responses for drawing 17 (tree on hill) show an interesting variation on the basic pattern. the default verb »he (‘exists’) appears to be able to fill either the adposition slot or the lpv/pred slot: (63) c&ha) pa»ha wa) »he »naz&i) tree hill a exists stands the tree is standing on a hill 23 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 24. (64) c&ha) pa»ha wa) el »naz&i) tree hill a there stands the tree is standing there on a hill (65) c&ha) pa»ha wa) a»ka)l »he tree hill a on top of exists the tree is there on top of a hill (66) c&ha) pa»ha wa) a»ka)l »naz&i) tree hill a on top of stands the tree is standing there on top of a hill (67) ?? c&ha) pa»ha wa) »naz&i) »he ?? tree hill a stands exists ?? the tree is standing on a hill in (63), »he (‘exists’) seems to fill the adposition slot, while in (65) it appears in its “normal” lpv/pred position (i.e. where it appears in all the other data). it is almost as if »he and »na)z&i have switched places in (63). i was unable to elicit a sentence like (67), so i do not know if the variation seen in (63) is simply another ordering possibility, or if there is something more going on. nevertheless, a possible explanation is that both terms are functioning as lpvs, except they now work in series (similar to the serial adpositions discussed above). another interesting anomaly in the data can be seen in the following sentences describing drawing 28. badwound interpreted the image as picture of woman on stamp17 and described it as follows: (68) »wi)ja) wa) »wiaskab el i»towapi (»ja)ke) woman a stamp there picture (sits) a picture of a woman sits on a stamp (69) * »wi)ja) wa) »wiaskab el i»towapi (»na)z&i) * woman a stamp there picture (stands) * a picture of a woman stands on a stamp (70) * »wi)ja) wa) »wiaskab el i»towapi (»ju)ke) * woman a stamp there picture (lies) * a picture of a woman lies on a stamp 17 badwound did not know the word for ‘stamp’ and therefore created one. 24 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 25 the first thing that caught my attention is that the lpv appears to be optional. this clearly challenges the notion that the only obligatory element in a lakhota sentence is a predicate. however, when i asked badwound to translate the lakhota phrase back to me in english, she said, “it is a picture of a woman that lies on the stamp.” this sounds like an embedded clause construction, so the mystery may simply be the result of a different type of construction. 3.2. results of lakhota data in relation to the universal conceptual categories hypothesis several previous studies have challenged the universal conceptual categories hypothesis at various levels of analysis, including implications that concepts like in and on may not be holistic primitives (brown 1994), that languages may partition the conceptual space in other ways which are learned just as early (bowerman 1996 and 2003), that precise (rather then general) axial geometry must often be expressed (levinson 1994), and that some languages code topological relations with (either completely or in combination with) contrastive locative verbs rather than adpositions (ameka & levinson 2003). the levinson and meira study approaches the debate by breaking down the overarching universal conceptual categories hypothesis into three progressively weaker hypotheses: hypothesis 1: all languages agree on basic categories like in, on, under, near, etc., in such a way that these notions form uniform, shared coremeanings for adpositions across languages. hypothesis 2: languages may disagree on the ‘cuts’ through this semantic space, but agree on the underlying organization of the space — that is, the conceptual space formed by topological notions is coherent, such that certain notions will have fixed neighborhood relations. hypothesis 3: the domain of topological relations constitutes a coherent semantic space with a number of strong attractors, that is, categories that languages will statistically tend to recognize even if some choose to ignore them. (levinson & meira 2003: 495-502) let’s see how the lakhota data compare to levinson and miera’s hypotheses. 3.2.1. hypothesis 1 this interpretation can be thought of as a “strong version” of the universal conceptual categories hypothesis. it includes what is sometimes referred to as prototype theory, where core concepts are considered to be universal but category boundaries can vary to some degree. 25 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 26. levinson and meira (2003) quickly refute this strong theory by comparing extensional maps for the multiple languages they studied. they found no evidence of prototype categories in the spatial topological domain. only a single grouping of three scenes was shared across all the languages they looked at: 53, 16, and 31 (under). similarly, the lakhota catagorizations seem to cut across boundaries identified for other languages. while lakhota does categorize 16 and 31 together (o»xlate), 53 was not part of the current study, so we cannot make any further inferences about a universal under category. 3.2.2. hypothesis 2 a weaker version of the universal conceptual categories hypothesis is based on comparative approaches that have been used in cross-linguistic exploration of basic color terms. for example, no language has been found that collapses purple and yellow into a single color category. instead, languages appear to build their color categories around foci that are “naturally salient” to human perception. cross-linguistic variations in color categories are explained by a combination of two factors: languages organize their categories around one or more of the six natural foci, and categories all have variable boundaries. levinson and meira state, “if the topological domain has a similar internal coherence, it should be possible to find a single fixed arrangement of the pictures such that those that are grouped together in one language remain contiguous even if they are separated by a category boundary in another language” (2003: 499). their best solution to such a fixed array is represented in figure 118, however it failed to meet the requirement for the languages they studied. no single arrangement could be found that did not result in some language-discontinuous categories. while this may seem to refute hypothesis 2, levinson and meira (2003) note that a perfect solution may eventually be found — unfortunately, it is an extremely difficult and computationally intensive problem to solve (71 factorial). therefore, before discounting this hypothesis, let’s see how the lakhota-specific extensional map shown in figure 2 compares with figure 1. as we can see, several lakhota categories seem to map well onto figure 1: ma»hel fits into the in region, i»wa)kab matches with over, and o»xlate fits within under. furthermore, i»sakhib provides a near match to near. the 18 figures are located in appendix b. 26 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 27 boundary scene (17) is acceptable in this hypothesis because it is still contiguous to the other scenes in the category. unfortunately, the remaining categories are not so clean. while the majority of a»ka)l fits nicely into the category of on, scenes (33) and (17) are dispersed across the space. a different kind of problem is presented by the two lakhota adpositions that code some kind of attachment (i.e. i»phaxlog, e»ta)). they both fall completely outside of the attached region, although along its edges. an interesting solution to this problem is to include three lpvs that also carry attachment connotations (i.e. o»tke, i»kwqke, i»jaskape) (see figure 3). now we can see that these mappings criss-cross the larger attached region. it is still discontinuous, but the overall impression is that all the attachment-related means of coding in lakhota seem to fall within the attached region. therefore, given the fixed array used by levinson and meira (2003), the lakhota data are inconsistent with hypothesis 2 — extensions of lakhota adpositions do not map onto the optimized fixed array of scenes in a continuous way. 3.2.3. hypothesis 3 the weakest of the three versions of the universal conceptual categories hypothesis may be the most elegant. if certain topological relations do in fact act as semantic “attractors,”19 we should be able to see clusters appear on a euclidean distance model (a two-dimensional representation of multidimensional space). levinson and miera claim to have found evidence for such clusters (see figure 4). however, they also state, “a crucial consideration is whether this particular pattern is an artifact of the particular languages we happen to have selected. that is a question that we cannot answer definitively — we can say only that the patterns now showing seem quite stable when further languages are added” (2003: 504). since levinson and miera’s clusters seem stable over a wide range of languages, it would be extremely interesting to see if the lakhota data follows the same pattern or not. unfortunately, the data analysis methodology required for 19 attractor is a term used in dynamical systems theory (complexity theory) to refer to a particular state towards which a complex system will tend to evolve, given enough time (see gleik, 1987; lorenz, 1996). 27 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 28. such an analysis (multidimensional scaling and cluster analysis) is beyond the scope of this project. we may, however, be able to get a general idea of how the lakhota data compares to levinson and miera’s cluster analysis by checking to see if any of the lakhota adpositions and/or lpvs fall outside of levinson and miera’s clusters.20 for example, figure 5 shows an enlargement of the attachment cluster region. of the 13 scenes that lakhota describes using attachment adpostions or lpvs, fully six of them are not represented in any of the clusters (i.e. 9, 20, 27, 30, 44, 41). it is possible that this may be due to including lpvs as well as adpositions in order to flesh out the attachment category. however, i believe that it instead shows that lakhota organizes its attachment category somewhat differently than the other languages represented in the cluster analysis. on the other hand, if we look at the two clusters that concern on (i.e. ontop and on-over), we find the lakhota organization largely follows the cluster analysis. lakhota uses a single adposition (a»ka)l ) for on, but its meaning can span three of the clusters. of the nine scenes for which a»ka)l is used, four fall into the on-over category, three into on-top, and two into attachment. levinson and miera state, “note that the ‘conflation’ of on/over suggests that on simpliciter is not a primitive (as on the orthodox view) but is composed of superposition plus or minus contact” (2003: 508). since the use of a»ka)l requires contact, it is broad enough in meaning to be used in all the scenes represented by the three different clusters. these two rough comparisons between the lakhota data and levinson and miera’s cluster analysis seem to point in different directions. some aspects of lakhota data may not fit into their model, while others seem to accord well. unfortunately, this study does not provide the kind of analysis that would be more conclusive. 20 a true cluster analysis of the lakhota data would plot each scene in the multidimensional space. any differences noted in this rough estimation may not in fact be significant (see section 4.1 for limitations of this study.). 28 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 29 4. discussion of results 4.1. limitations of the study clearly, a study of this size (utilizing only a single consultant) has many limitations. first and foremost, generalizing from a specific individual’s utterances to the structure of a language is a significant leap of faith. without comparing this dataset to responses from other native speakers, we cannot know to what extent badwound’s patterns indicate preferences of the language itself or are instead idiosyncratic. further study utilizing a larger subject pool and statistical methodologies will eventually be needed to fully understand the extent of any variability. another major limitation of this project concerns cross-linguistic studies in general. research designs attempting to understand something about “universal” conceptual categories may be inherently problematic. various types of researcher bias (e.g. linguistic, cultural, philosophical) are difficult if not impossible to avoid, even when they are acknowledged. furthermore, human cognition is a socially embodied process rather than the abstract, disembodied form of “reason” proposed by most semantic typologists (hutchins 1995). therefore, attempting to abstract purely conceptual information from decontextualized drawings may provide a skewed picture (see goodwin 1997 for an excellent critique of basic color term studies.) in other words, it is possible that the very methodology used in this project is the linguistic equivalent of forcing a square peg into a round hole. finally, i cannot offer any reliable analysis or conclusion for the comparison between the lakhota data and levinson and miera’s hypothesis 3 because the kind of data analysis required is beyond the scope of this project. given these limitations, there are still some observations that we can make. 4.2. conclusion although languages use different means to express spatial topological relations, it has been generally assumed that the underlying conceptual categories are universal. consequently, it was thought that individual languages simply map their particular coding means onto the underlying categories. recent crosslinguistic research in the spatial topological domain, however, has challenged the assumptions of this universal conceptual categories hypothesis on a variety of levels. this project has analyzed and described the basic locative construction in lakhota. it has explored each of the elements in detail, identified which are optional and which are obligatory, noted the acceptable combinations and orderings of elements, and identified multiple animacy criteria. since lakhota is a language that has called into question universal theories in other areas of 29 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 30. linguistics, a comparison was made between the lakhota data and the conclusions from levinson and meira’s groundbreaking cross-linguistic study (2003). levinson and meira break the universal conceptual categories hypothesis down into successively weaker hypotheses, and conclude that there is only evidence for the third and weakest version. the lakhota data analyzed here confirms that hypothesis 1 (strong view) is unfounded — there is no evidence of prototype categories in the spatial topological domain. the lakhota data also confirms that hypothesis 2 is unsupportable. extensions of lakhota adpositions do not map onto the optimized fixed array of scenes in a continuous way. therefore, languages neither agree on the ‘cuts’ through semantic space, nor their underlying organization. finally, lakhota provides some evidence to challenge levinson and meira’s proposed solution, hypothesis 3 (strong attractors in semantic space create a tendency for languages to code certain categories in certain ways). consequently, we may be forced to formulate an even weaker version of the universal conceptual categories hypothesis, or instead, simply discount the idea of universal categories in the domain of spatial topological relations. 30 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 31 references ameka, felix and stephen levinson. 2003. “positional and postural verbs.” in special issue of linguistics, to appear. ayano, seiki. 2001. the layered internal structure and the external syntax of pp. durham: university of durham dissertation. baker, mark. 2003. lexical categories: verbs, nouns, and adjectives. cambridge: cambridge university press. bowerman, melissa. 1996. “learning how to structure space for language: a cross-linguistic perspective.” in language and space, bloom, peterson, nadel, and garrett (eds.), 385–436. cambridge, ma: mit press. bowerman, melissa. 2003. “space under construction: language-specific spatial categorization in first language acquisition.” in language in mind, gentner and goldin-meadow eds., 387–427. cambridge, ma: mit press. bowerman, melissa and soonja choi. 2001. “shaping meanings for language: universal and language-specific in the acquisition of spatial semantic categories.” in language acquisition and conceptual development, bowerman and levinson (eds.), 475-511. cambridge: cambridge university press. bowerman, melissa and eric pedersen. 2003. cross-linguistic perspectives on topological spatial relationships. eugene: university of oregon, and nijmegen: max planck institute for psycholinguistics, ms. brown, penelope. 1994. “the ins and ons of tzeltal locative expressions: the semantics of stative descriptions of location.” linguistics, 32 (4/5): 743–90. gleick, james. 1987. chaos: making a new science. new york: penguin books. goodwin, charles. 1997. “the blackness of black: color categories as situated practice.” in discourse, tools and reasoning: essays on situated cognition, b. burge (ed.), 111--140. new york: springer-verlag. hutchins, edwin. 1995. cognition in the wild. cambridge, ma: mit press. ingham, bruce. 2003. lakota: languages of the world/materials 426. germany: lincom europa. kay, paul and luisa maffi. 1999. “color appearance and the emergence and evolution of basic color lexicons.” american anthroplogist, 101: 743-760. lemmens , maarten. (2006). “caused posture: experiential patterns emerging from corpus research.” in stefan gries and anatol stefanowitsch (eds.) corpora in cognitive linguistics. vol. ii: the syntax-lexis interface. berlin: mouton de gruyter. levinson, stephen. 1994. “vision, shape and linguistic description: tzeltal bodypart terminology and object description.” in john b. haviland and stephen c. levinson (eds.) space in mayan languages (special issue of linguistics 32.4), 791–855. berlin: mouton de gruyter. 31 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 32. levinson, stephen and sergio meira, s. 2003. “‘natural concepts’ in the spatial topological domain — adpositional meanings in cross-linguistic perspective: an exercise in semantic typology.” language, 79 (3): 485-516. lorenz, edward. 1996. the essence of chaos. seattle: university of washington press. rood, david and allan r. taylor. 1996. “sketch of lakhota, a siouan language.” in handbook of north american indians. washington dc: smithsonian institution. vol. 17, 440-482. rood, david. 2003. “two lakhota locatives and the role of introspection in linguistic analysis.” in motion, direction, and location in language, seibert and shay (eds.), 255-58. amsterdam: john benjamins. van valin, robert, d. 2001. an introduction to syntax. cambridge: cambridge university press. 32 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 33 appendix a – subset of bowerman’s topological relations picture series 33 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 34. 34 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 35 appendix b – figures figure 1. notional areas in fixed array (from levinson & meira 2003: 502). 35 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 36. figure 2. lakhota adpositions mapped onto fixed array. 36 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 locative constructions in lakhota 37 figure 3. lakhota adpositions and lpvs of attachment mapped onto fixed array. 37 sikos: locative constructions in lakhota published by cu scholar, 2010 colorado research in linguistics, volume 22 (2010) 38. figure 4. alscal plot for tiriyó, yélî, dnye, ewe, lavukaleve, trumai, yukatek, lao, dutch, and basque (from levinson & meira 2003: 505). figure 5. enlargement of attachment cluster region (from levinson & meira 2003: 508). 38 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/5 doi: https://doi.org/10.25810/06db-m579 colorado research in linguistics 6-2010 locative constructions in lakhota: evidence for/against “universal conceptual categories” in spatial topology les sikos recommended citation cril_sikos_v5 a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) tao lin university of colorado boulder this research investigates the linguistic features that can predict native speakers’ choice of three word order alternations of one kind of verb-direction constructions (vdcs) in mandarin chinese, i.e., object initial, middle, and final types. previous work on the contributing features is done through qualitative observations, and the direction of some predictions is unclear. this paper provides corpus based quantitative studies on both monofactorial and multifactorial analyses, which clarify the effectiveness of most of these parameters. in addition to the discussed features, some new features, such as the use of certain classes of verbs and imperatives, are also introduced in order to enhance the model; further directions and improvements are also. keywords: multinomial logistic regression, word order, constructions, chinese 1. introduction like many languages, such as english and arabic, word order variation in chinese is a common device in syntactic alternations. english verbal particle constructions (vpcs) can allow different placements of the verbal particles, which are realized in prepositions (dehé 2002, jackendoff 2002 & wasow 2002). (1) a. john took out a book. b. john took a book out. in the chinese translation of sentence 1, instead of using the preposition ‘out,’ chinese can take two verbal particles: chu ‘exit’ and lai ‘come’ marked as particle 1 and 2 (p1 and p2) in the examples. their origins all came from directional verbs (lamarre 2007). there are three possible alternations. (2) a. (object initial type) ta na yi ben shu chu lai he take one cl book exit(p1) come(p2) 1 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 b. (object middle type) ta na chu yi ben shu lai he take exit (p1) one cl book come(p2) c. (object final type) ta na chu lai yi ben shu he take exit(p1) come(p2) one cl book ‘he took out a book.’ these three alternations share the same meaning in example 2. the only difference is the placement of the object ‘a book.’ this study aims at adopting a quantitative approach to discover how native chinese speakers prefer one alternation over another. in chinese there are 28 directional particles (liu 1998). however, in this research, i focus only on one combination (chu lai), because the linguistic factors related to it different orders have been well studied. to further narrow down the searching scope, i only choose one verb, na ‘to take,’ as the target verb. next, i will introduce some linguistic factors that are believed to affect the choice of alternations, and use it to make up the first hypothesis (model a). i collect the data in section 3. based on my observation on the corpora data, i discuss more possible factors in section 4. i build up the second hypothesis (model b) by adding these factors to model a, which contains a set of hypotheses to be tested in factorial analysis. i identify the statistical techniques in section 5. the results in section 6 suggest that model b is statistically well supported. while the results have many implications for future work, expanding to other kinds of vdc-related phenomena would be an obvious next step. 2. variables that purportedly govern the alternations a lot of work on the syntactic alternations of vdcs has been done within the last twenty years. mostly using introspective data, researchers believe that the emergence of the three alternations can result from features in syntactic (2.3 & 2.4), semantic (2.2), phonological (2.1), and discursive (2.5 & 2.6) domains. 2.1. the length of noun phrases (nps) sun (2012) argues that the syllable/character length is related to the use of the three alternations. the object-initial alternation tends to exclude long objects, and the object-final ones are less likely to have short objects. 2 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 (3) a. na chu banfa lai take go.out (p1) method come (p2) ‘to come up with a plan’ b. na chu lai yi ge feichang bianjie de caomiaoyi take p1 p2 one cl very convenient assoc scan.machine ‘to take out a very user-friendly scanner’ the bolded parts are the object nps. the object-middle example 3a contains a two-syllable np, while the object-final example 3b has ten syllables. 2.2. concreteness of nps when the object is an abstract item, the object-middle alternations are more common (liu 1998). in this condition, an object-final alternation can usually be replaced by an object-middle one, but it is not true for an object-initial alternation. (4) men-feng-li chuan chu lai yi zhen ge-sheng door-gap-inside pass p1 p2 one cl song-sound ‘sounds of music came from the gap in the door.’ in example 4, it is grammatical to transform the object-final alternation into an object-middle one by putting the object ‘a song’ between particle 1 and 2. however, the placement does not work for an object-initial alternation. 2.3. definiteness of nps in chinese, a definite np can be indicated by demonstratives (i.e., zhe ‘this,’ or na ‘that’), proper nouns as genitive (i.e., fuermosi de shu ‘holmes’ book’), and personal pronouns (i.e., ta de ‘his’). compared to the non-object-final alternations, nps in object-final ones are less likely to be definite (zhang 1991 & niu 2002). my pilot search on google shows that there are only 12 results for the object-final alternation na chu lai zhe ben shu ‘to take out this book,’ while the other two alternations with demonstratives zhe ‘this’ have much more results. 2.4. the aspect marker le after the verb the influence of the post-verb aspect marker le on the choice of the alternations is controversial. based on introspection, lu (2002) concludes that there is no alternation preference 3 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 for the aspect marker. chen (1982) believes that le only appears in the object-initial alternation. zhang (1991) argues that the other two alternations can also have le. his corpus study shows that the object-initial alternation tends to have the post-verb aspect marker le and that the other alternations have much less use of le. (5) ta na le zhi chu lai he take asp paper p1 p2 ‘he takes paper out.’ the object-initial type in example 5 can be replaced by an object-middle type, but cannot be replaced by an object-final one. 2.5. old and new information of nps if an np is definite or has a referent in the previous clauses, then i call it old information. an np that has never been introduced before is new information. the object-final and objectmiddle alternations tend to introduce new information (zhang 1991). (6) wo yiwei ta hui na chu yi feng xialüdi de xin lai, i think she will take p1 one cl charlotte assoc letter p2 danshi bingbu jian ta na xin chu lai but neg see she take letter p1 p2 ‘i thought she would take out charlotte’s letter, but i didn’t see that she took out any.’ in example 6, there are two verb-direction constructions with the object xin ‘letter.’ although the second ‘letter’ is indefinite, it is old information because of its prior mention. 2.6. zero-anaphora zero-anaphora means a gap in a clause whose referent is the object of a vdc in the previous clause. object-final alternation is more highly correlated to the use of zeroanaphoras (sun 2012), as in 7. 4 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 (7) fuermosi na chu lai yi xiao pian qianbi mu-xie, holmes take p1 p2 one small cl pencil wood-dust, shangmian you zimu ‘n’ top have alphabet ‘n’ ‘holmes takes out a small piece of pencil shaving. it has the alphabet ‘n’ at the top.’ in the second clause, there is no subject, but one can recover the omitted subject by using the object in an object-final alternation (the bolded part). to summarize section 2, i synthesize model a: the alternation choice of verbdirection construction depends on length, concreteness, definiteness, and information type of an object np, the use of aspect marker le, and zero-anaphora. 3. data collection this study aims at measuring how linguistic variables of interest can contribute to the alternation choice for vdcs. 350 tokens were collected and annotated with linguistic factors for each alternation. only one verb na ‘to take’ and one direction combination chu lai ‘come out’ are searched in the corpora. i planned to use chinese treebank 6.0 (ctb) (xue et al. 2005) and chinese gigaword (huang 2009) as the main source of corpora. however, they did not have enough data for this specific search task. for example, i found only 20 tokens of object-middle alternation in chinese gigaword, and no result of object-initial alternation in ctb. in order to find plenty of instances, other modern chinese corpora/search engines were used, as in table 1. table 1. corpora and search engines used in this study name size genre corpus in center of chinese linguistics in beijing university (ccl) (zhan et al. 2003) 477 million characters written chinese national broadcast language media resource 241 million characters spoken chinese national corpus (xiao 2012) 19 million characters written central china normal university corpus 44 million written corpus zhtenten11 (jakubíček 2013) 1 billion words written google, baidu, & sougou 5 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 the collected 1,050 tokens came from different genres and formats, and are not balanced corpus data. for example, some tokens do not have a minimal context for the discursive annotation, which may have influence on the analysis of features in section 2.5 and 2.6. 4. revised set of variables based on my observation, five more factors complementary to those in section 2 are considered in my hypothesis. i augment the original six factors in model a by adding the following five new factors (4.1-4.5), and in so doing i tailor model a to model b. 4.1. location/source (thematic role) according to the argument structure of the verb ‘to take,’ the typical thematic role assignment is ‘agent take p1 p2 theme.’ however, other thematic roles, such as location and source, also appear in the alternations. the expressions of location or source are usually featured in the preposition cong ‘from,’ postposition li/limian/zhong ‘inside,’ or both. when p1 and p2 are not adjacent (the object-middle alternation), location and source are less likely to appear. 4.2. quantification the object can be modified by quantifying expressions, such as classifiers (i.e., ge ‘the general classifier’), quantifiers (i.e., xuduo ‘many’ or yixie ‘some’), and numbers, or the object itself is conceptually a quantity such as certain amounts of money (i.e., ji qian wan yuan ‘ten million yuan’). in our data, the object-final alternation is most likely be quantified. 4.3. imperatives imperatives are ‘verb forms or construction types that are used to directly command the addressee to perform some action’ (payne 1997: 303). lü (1992) predicts that only the object-final type prevents imperatives, while the other two can have imperative functions. if there is no imperative for objective-final alternation in the data, then it is likely that this statement is true. 4.4. modal verbs modal verbs are (auxiliary) verbs that can deal with judgments and evidentials, and so on (palmer 2001: 100). like english, chinese also has a relatively close set of modal verbs, such as ying/gai/yinggai/yingdang ‘should,’ bixu ‘must,’ and neng/nenggou ‘can,’ as in 8. 6 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 (8) taiwan dang-ju yinggai na chu shiji xingdong lai taiwan current-government should take p1 practical action p2 ‘taiwan government should take out practical action.’ it is possible that object-middle alternation is more correlated to modal verbs than the other two types. i use the wordlists from peng 2005 and fan et al. 1987 as a matching guideline to find possible modal verbs. 4.5. request verbs like modal verbs, if there is a verb for request in the main clause, then the objectmiddle alternation is more likely to appear. there are two semantic sub-classes for request: verbal and non-verbal. if an imperative speech act is realized in the complement clause, a verbal request in the main clause is expected, such as ‘to order’ and ‘to require.’ (9) ta-men yaoqiu pingguo na chu yixie shuju lai 3p-pl require apple take p1 some data p2 ‘they require that apple take out some data.’ the non-verbal request class includes verbs like ‘to hope’ and ‘to expect,’ in which a speaker has a will to change other people’s action. the english translation of chinese request verbs is roughly mapped to classes order-60-1, urge-58.1, and wish-62 in english verbnet (kipper et al. 2008). 5. hypotheses, variable assignment and statistical techniques the hypotheses of this study is based on the eleven predictions in the alternation preference in model b. compared to model a which has fewer parameters, i hope model b can have a better explanation on the variation of the 1,050 tokens in our data. before any factorial analysis is performed, i need to translate various linguistic factors into variable types and values that are computable. 7 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 table 2. variable assignment variables variable type value alternation type nominal 1=‘obj-initial,’ 2 =‘obj-middle,’ and 3=‘obj-final’ syllable length of object interval any positive integer concreteness of object nominal 0=‘concrete,’ 1=‘abstract’ definiteness of object nominal 0=‘indefinite,’ 1=‘definite’ aspect marker nominal 0=‘no marker le,’ 1=‘marker le’ information structure nominal 0=‘old,’ 1=‘new’ zero-anaphora nominal 0=‘no,’ 1=‘yes’ thematic role nominal 1=‘source,’ 2=‘location,’ and 3=‘neither’ quantifying expression nominal 0=‘no,’ 1=‘yes’ imperative nominal 0=‘no,’ 1=‘yes’ modal verb nominal 0=‘no,’ 1=‘yes’ request verb nominal 0=‘no,’ 1=‘yes’ after the collected tokens were annotated with standards in table 2, i use ibm spss software to make monofactorial and multifactorial analyses (morgen et al. 2012; sweet et al. 1999). the correlation coefficients are lambda and eta, which separately deal with the relationship between nominal variables and the alternation types, and that between nominal variables and interval variables. multinomial logistic regression analysis will tell us how one factor value predicts the choice of an alternation type. 6. results 6.1. descriptive statistics table 3 includes the counts and their percentage in the sample tokens. i tested whether our sub-hypotheses (2.1-2.6 & 4.1-4.5) are well indicated in the counts and recorded the results in the last column. it is obvious that the function of zero-anaphora is rejected and that imperatives do not apply to the object-final alternation. however, counts are easily affected by extreme values, and cannot tell us any correlationship, which makes monofactorial and multifactorial methods necessary. 8 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 table 3. basic counts i obj-initial obj-middle obj-final good match? thematic roles source (1) 4%(13/350) 1%(4/350) 5%(18/350) y location (2) 11%(38/350) 3%(11/350) 7%(25/350) y abstract obj (1) 15%(54/350) 63%(222/350) 8%(27/350) y definite obj (1) 7%(23/350) 4%(14/350) 5%(18/350) ? asp marker (1) 15%(51/350) 0%(1/350) 1%(4/350) y old information (1) 27%(94/350) 3%(12/350) 11%(38/350) ? zero anaphora (1) 9%(31/343) 3%(9/348) 4%(13/348) n quantifying expression (1) 61%(212/350) 34%(120/350) 87%(306/350) y imperative (1) 8%(28/350) 5%(19/350) 0%(0/350) y modal verb (1) 15%(52/350) 47%(163/350) 2%(8/350) y request verb (1) 7%(24/350) 22%(76/350) 0%(2/350) y counts for the length of objects alternations is in table 4. table 4. basic counts ii (syllable length of objects) alternation 1 to 5 6 to 10 11 to 15 more than 16 total obj-initial 326 23 0 1 350 obj-middle 259 70 17 4 350 obj-final 248 81 15 6 350 total 833 174 32 11 1050 the distribution pattern supports sun’s observation on the influence of the object length in section 2.1. almost all the objects for object-initial alternation (349 out of 350) have less than five syllables, while object-middle and final alternations can allow some objects longer than 11 syllables. 6.2. monofactorial result the first monofactorial technique i use is lambda, which computes the correlation between any of the nominal linguistic factors and the alternation type. pr(>|z|) shows whether the correlation is significant. if it is less than 0.001or 0.01, the correlation is reliable. table 3 indicates that except for thematic roles (location and source) and definiteness, other linguistic factors do have a low correlation with the choice of alternations. the second monofactorial technique is eta, which is not included in table 5. the eta value for syllable length is 0.35, which means that there is an intermediate correlation between the syllable length of objects and the alternation choice. 9 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 table 5. results of monofactorial analysis lambda pr(>|z|) significance good match? abstract obj 0.279 0.000 ***(<0.001) y quantifying expression 0.266 0.000 ***(<0.001) y old information 0.122 0.000 ***(<0.001) y request verb 0.106 0.008 **(<0.01) y asp marker 0.072 0.000 ***(<0.001) y thematic roles 0.059 0.111 ns n modal verb 0.037 0.000 ***(<0.001) y zero anaphora 0.032 0.000 ***(<0.001) y definite obj 0.011 0.756 ns n although monofactorial analysis establishes a steady relationship between alternation choice and linguistic factors, it has weaknesses too. first, lambda only works for any single factor and alternation preference. that being said, it is likely that the interrelationship between two factors (i.e., the correlation between concreteness and quantifying expressions) may be underestimated (gries 2001 & 2003). since abstract nouns rarely combine with quantifying expressions, the correlation between quantifying nouns and alternation choice may not be reliable. i need a multifactorial method that can rule out the effect of possible meditators and support a cognitively realistic account of alternations. 6.3. multifactorial result following the tradition of using regression models (bresnan et al. 2007), multinomial logistic regression (mnr) is used as a multifactorial method in this study, and shows the logistic coefficient between each linguistic factor (independent variable) and alternation preference (dependent variable) in an interrelation. mnr usually assumes ‘one outcome category as a default case against which other categories are contrasted’ (mari-sanna et al. 2012: 5). all significant results are shown in table 6. 10 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 table 6. results of the multifactorial analysis reference dependent variable independent variable weight pr(>|z|) odds ratio c a information (0) 1.567 <0.001*** 4.794 quantifying expression (0) 1 <0.001*** 2.717 length of np -0.384 <0.001*** 0.681 modal verb(0) -2.477 <0.001*** 0.084 request_verb(0) -2.53 <0.01** 0.08 asp marker (0) -3.137 <0.001*** 0.043 imperative(0) -17.68 <0.001*** 2.10e-08 b quantifying expression (0) 2.09 <0.001*** 8.082 request_verb(0) -0.395 <0.001*** 0.019 concreteness(0) -1.724 <0.001*** 0.178 modal verb(0) -3.776 <0.001*** 0.023 b a information (0) 2.064 <0.001*** 7.878 request_verb(0) 1.419 <0.001*** 4.135 modal verb(0) 1.299 <0.001*** 3.666 concreteness(0) 1.248 <0.001*** 3.483 quantifying expression (0) -1.09 <0.001*** 0.336 asp marker (0) -2.763 <0.01** 0.063 length of np -2.95 <0.001*** 0.745 c request_verb(0) 3.95 <0.001*** 51.928 modal verb(0) 3.776 <0.001*** 43.632 concreteness(0) 1.724 <0.001*** 5.61 length of np 0.089 <0.05* 1.093 quantifying expression (0) -2.09 <0.001*** 0.124 a b asp marker (0) 2.763 <0.01** 15.852 information (0) 2.064 <0.001*** 1.27e-01 quantifying expression (0) 1.09 <0.001*** 2.975 length of np 0.295 <0.001*** 1.343 concreteness(0) -1.248 <0.001*** 0.287 modal verb(0) -1.299 <0.001*** 0.273 request_verb(0) -1.419 <0.001*** 0.242 c asp marker (0) 3.137 <0.001*** 23.033 request_verb(0) 2.53 <0.01** 12.558 modal verb(0) 2.477 <0.001*** 11.902 length of np 0.384 <0.001*** 1.468 quantifying expression (0) -1 <0.001*** 0.368 information (0) -1.567 <0.001*** -0.209 11 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 the ‘reference’ contains different alternations as a baseline for comparison. for example, ‘reference = c’ and ‘dependent variable = a’ mean that relative to c, a has a certain tendency. also, in this table, if the odd ratio is bigger than one, then the weight should a positive number, which tells us the dependent variable has a positive effect on the alternation choice. otherwise there would be a negative effect. the full interpretation of table 6 is shown in section 7 as the conclusion. 7. conclusion this study first introduces model a (including eleven factors), and improves model a into model b by adding five more linguistic features. the basic aim is to use factorial analysis to testify the hypotheses in model b. the monofactorial analysis shows that when there is no variable interrelationship, all linguistic factors correlate with the alternation preference, except for definiteness and the co-occurrence of location and source. however, multifactorial analysis gives us much more information about the weight of linguistic variables when any two of the three alternations are compared. the influence of zero-anaphora is rejected. based on table 6, i can draw a conclusive sketch for this study. first, the object-initial and middle alternations are compared relative to the object-final alternation. the distinguishing factors of an object-initial alternation are: a shorter object, the aspect marker le, new information in the object, imperative, no quantifying expression, modal verbs and request verbs. the distinguishing factors of an object-middle alternation are an abstract object, no quantifying expression, modal verbs, and request verbs. second, the object-initial and final alternations are compared relative to the object-middle alternation. the distinguishing factors of an object-initial alternation are: a shorter and more concrete object, the aspect marker le, old information in the object, quantifying expression, and no modal verbs or request verbs. the distinguishing factors of an object-final alternation are longer and more concrete objects, quantifying expressions, and no modal verb or request verb. third, the object-middle and final alternations are compared relative to the object-initial alternation. the distinguishing factors of object-middle alternation have longer and more abstract object objects, no aspect marker or quantifying expression, old information in the object, modal verb, and request verb. the distinguishing factors of an object-final alternation are a longer 12 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 object, the aspect marker le, new information in the object, quantifying expressions, modal verb, and request verb. in general, i find that most linguistic factors in model b (eight out of eleven) can help us discriminate the three verb-direction alternations. 8. further directions the main contribution of this study is to firstly use statistical correlation and regression methods to discover how each syntactic, semantic, discursive, or phonological factor contributes to the choice of a verb-direction alternation, which is commonly believed to be interchangeable among three alternations. there are several directions i can improve this study and apply the result to other linguistic studies. first, i plan to apply the same method to other types/senses of vdcs. according to liu (1998), there are other directional verbs which involve multiple motion types such as ‘raising,’ ‘moving into,’ and ‘moving up/down.’ object displacement can work for some of them. also, vdcs can not only encode special movement, they can also express different resultative and aspectual interpretations. it is highly likely that even for the same direction particle(s), different meanings will bring different results in factorial analyses. second, the role of verb semantics in this study is not fully discovered. in order to get an indepth analysis of alternation choice, i just choose one verb na ‘to get’ for neutralization. it is well known that different verbs conceptually evoke different argument structures. for example, the chinese verb zou ‘to walk’ can also combine with p1 and p2, but the thematic role of its object can be an agent, which na ‘to get’ can never do. more verbs need to be collected in order to testify the effect of thematic roles in alternation choice. also, semantic classification of verbs remains another question. for example, i can compare the factorial analysis results among verbs in various unrelated classes, using verbnet. also, there exists an interrelationship between verb’s semantic classes. synsets with prototypical verbs like ‘to hope’ or ‘to ask’ do have a restriction on the alternations. these classes are positively correlated to the emergence of an object-middle alternation. it is still unknown how this interrelationship can improve the designing work of a computational lexicon, such as chinese verbnet. third, i disprove three factors in model b, resulting in the development of model c. model c can be further used to discriminate more alternations related to vdcs, such as those carrying 13 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 passive marker bei and disposal marker ba. more interestingly, in our search of the ccl corpus, there is a new alternation whose form is ‘verb p1 object p2 p3,’ which appeared in low frequency and was not discussed by the previous references. it is likely that this alternation will be acquired and used by more chinese speakers. factorial analysis will help us discover how one language evolves by looking at the change in alternation choice. finally, from the perspective of natural language processing (nlp), our statistical method in multifactorial analysis needs to be improved. the output of regression in spss is a comparison between two alternations. this ‘two against one’ technique should be changed so it can be compatible with the standard multi-class classification in nlp. additionally, the frequency distribution of vdcs is equally important as the coefficients in this study. references chen, xinchun. 1982. tong fuhe buyu bingjian de quxiang buyu de weizhi (the position of object in co-occurrence with compounding direction particles), newsletter of chinese language 5. dehé, nicole. 2002. particle verbs in english: syntax, information structure and intonation (vol. 59). john benjamins publishing. fan, xiao; du, gaoyin; and chen, guanglei. 1987. han yu dong ci gai shu (an investigation on chinese verbs). shanghai: shanghai education press. huang, chu-ren. 2009. tagged chinese gigaword version 2.0, ldc2009t14. linguistic data consortium. gries, stefan. 2001. a multifactorial analysis of syntactic variation: particle movement revisited. journal of quantitative linguistics 8(1). 33-50. gries, stefan. 2003. multi-factorial analysis in corpus linguistics. continuum international publishing group. jackendoff, ray. 2002. english particle constructions, the lexicon, and the autonomy of syntax. verb-particle explorations. 67-94. kipper, karin; anna korhonen; neville ryant; and martha palmer. 2008. a large-scale classification of english verbs. language resources and evaluation journal 42(1). 21-40. lamarre, christine. 2007. the linguistic encoding of motion events in chinese: with reference to cross-dialectal variation. typological studies of the linguistic expression of motion events 1. 3-33. 14 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 liu, yuehua. 1998. quxiang buyu tongshi (chinese directional particles: an investigation). beijing: beijing language and culture university press. lu, jianming. 2002. dongci hou quxiang buyu de weizhi wenti (on the positions of the directional particles after the verbs). chinese teaching in the world 1.5-17. lü, shuxiang. 1992. tongguo duibi yanjiu yufa (using comparison to study chinese grammar). chinese language teaching and research 2.4-18. jakubíček, m., kilgarriff, a., kovář, v., rychlý, p. and suchomel, v., 2013, july. the tenten corpus family. in 7th international corpus linguistics conference cl. 125-127. mari-sanna, paukkeri; väyrynen, paukkeri,; and arppe, antti. 2012. exploring extensive linguistic features sets in near-synonym lexical choice. computational linguistics and intelligent text processing: 13th international conference, cicling 2012, new delhi, india, march 11-17, 2012, proceedings ii. 1-12 morgan, george a., nancy l. leech, and karen c. barrett. 2012. ibm spss for introductory statistics: use and interpretation. routledge. niu, yanmin. 2002. a word order study on chinese verb-direction construction. (master's thesis). china capital normal university (beijing, china). palmer, frank robert. 2001. mood and modality. cambridge university press. payne, thomas e., and thomas edward payne. 1997. describing morphosyntax: a guide for field linguists. cambridge university press. peng, lizhen. 2005. on modality of modern chinese. (doctoral dissertation). fudan university (shanghai, china). sun, yanling. 2012. a study of order of compound directional complements and objects in teaching chinese as a second language. (master's thesis). beijing university. sweet, stephen a., and karen grace-martin. 1999. data analysis with spss. allyn & bacon. wasow, thomas. 2002. postverbal behavior. center for the study of language and inf. xiao, hang. 2012. guojia yuwei xiandai hanyu yuliaoku jieshao (an introduction of modern chinese corpus by the institute of applied linguistics of ministry of education in prc). online: http://corpus.zhonghuayuwen.org/resources/corpusintroduction2012.pdf xue, naiwen, fei xia, fu-dong chiou, and marta palmer. 2005. ‘the penn chinese treebank: phrase structure annotation of a large corpus.’ natural language engineering 11 (2). 207-238. zhan, weidong., guo, rui., chen, yirong., 2003. the ccl corpus of chinese texts: 700 million 15 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 chinese characters, the 11th century b.c. present, available online at the website of center for chinese linguistics (abbreviated as ccl) of peking university. online: http://ccl.pku.edu.cn:8080/ccl_corpus zhang, bojiang. 1991. dongqushi li binyu weizhi de zhiyue yinsu (on the restricting factors of objects in chinese verb-direction constructions). chinese learning 6. 4-8. 16 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/3 doi: http://dx.doi.org/10.33011/cril.24.1.3 appendix – abbreviations 3p: the third person asp: aspect assoc: associative de cl: classifier neg: negation pl: plural pn: the n-th directional particle 17 lin: a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) published by cu scholar, 2019 colorado research in linguistics 6-2019 a multifactorial analysis on the syntactic variations in one kind of chinese verb-direction construction (vdc) tao lin recommended citation microsoft word lin-cril2019-final.docx polysemy in the mental lexicon colorado research in linguistics. june 2008. vol. 21. boulder: university of colorado. © 2008 by susan windisch brown. polysemy in the mental lexicon susan windisch brown university of colorado the semantic ambiguity of lexical forms is pervasive: many, if not most, words have multiple meanings. for example, one can draw a gun, draw water from a well, or draw a diagram. despite the frequency of this phenomenon, how human beings store and access these meanings is an open question. do we have a separate representation in our mental lexicon for each “sense,” or do we store only one very generalized or core meaning for each word? if the latter, do we generate the nuances of each separate sense by rule or by accessing subrepresentations? to even speak of senses in this way implies that we can clearly identify the separate senses of a word. in this study, we use priming in a semantic decision task to investigate the effect of different levels of meaning relatedness on language processing. both response time and accuracy followed a linear progression through four categories of meaning relatedness. these results suggest that the distinction between a single phonological form with unrelated meanings (homonyms) and a single form with related meanings (polysemes) may be more one of degree than of kind. they also imply that related word “senses” may be part of a continuum or cluster of meanings rather than discrete entities. in addition, results from specific comparisons between groups do not support the theory that each sense of a word has an entirely separate mental representation. 1. introduction pinning down how people store and process words with multiple meanings has proven difficult to do, and the wide variety of theories currently in contention illustrates how little consensus has been reached. at one end of the spectrum lie theories in which each phonological form is connected to one complex semantic representation, with precise senses of a word only realized in context (kintsch 2001, 2007; ruhl 1989). in these theories, the unrelated meanings of homonyms coexist in the single semantic representation and are resolved in context by the surrounding words in the utterance. at the other end lies the theory that each sense of the same form, whether unrelated (homonyms) or related (polysemes), i would like to thank my advisors on this project, martha palmer and walter kintsch, for their support and assistance in designing and executing this project, and their helpful comments on an earlier draft of this paper. i would also like to thank al kim for his assistance. i gratefully acknowledge the support of the national science foundation grant nsf0415923, word sense disambiguation. susan windisch brown 1 windisch brown: polysemy in the mental lexicon published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 2 has its own semantic representation, with nothing but the phonological form in common (klein & murphy 2001). in the middle of the spectrum lie several popular theories, all of which suggest that related senses share a general or core semantic representation. for some, this core is the only thing stored in long-term memory. individual senses are generated online through some combination of pragmatics, general patterns of extension, and reasoning, given the context of the word (nunberg 1979; pustejovsky 1995). for others, individual senses may be stored, but related ones share a core representation or occupy overlapping areas of semantic space (klepousniotou & baum 2006; pylkkänen, llinás & murphy 2006; rodd, gaskell & marslenwilson 2002, 2004). these theories can be compared to many dictionaries, in which homonyms have separate entries but polysemes are listed as a single entry, with subentries for each sense. many researchers have investigated this question with lexical decision tasks, in which subjects decide whether a word they are seeing is a real word. response times are thought to indicate how easily a subject has accessed the word. several such studies were designed to compare homonymy and polysemy. azuma and van orden (1997) used words rated for the number of their senses and the relatedness of the senses. they found that, after controlling for number of senses, words with related senses had faster reaction times than words with unrelated senses, suggesting that polysemy facilitated lexical access as compared to homonymy. other studies suggested that polysemous words are identified in lexical decision tasks more quickly than single-meaning words (beretta, fiorentino, & poeppel, 2005; klepousniotou 2002; rodd, gaskell, & marslen-wilson, 2002, 2004). in addition, rodd, gaskell and marslen-wilson (2002) found the opposite effect for homonyms; namely, that homonymy slows response times. these two effects, the authors argue, reflect different processing methods for homonyms and polysemes, which stem from different mental storage architectures. the slower response times for homonyms are attributed to competition between separate mental “entries” or representations, whereas the faster response times for polysemous words are attributed to the ease of accessing a single, broad mental representation. a different pattern emerged, however, when subjects were forced to make judgments about meaning. klein and murphy (2001) used a “sensicality” decision task, in which subjects were asked to judge whether short phrases “made sense” or not. subjects saw one phrase, such as “daily paper”, then saw another phrase with the same noun, either with the same meaning, “liberal paper”, or with a different meaning, “lined paper”. (nonsense foils, such as “angry plate”, were interspersed with the target pairs.) the authors found that with this task, reaction time differences between homonyms and polysemes disappeared. that is, after processing one meaning of a word, the processing of another related meaning was inhibited rather than facilitated, compared to processing the same meaning again. the inhibition resulting from unrelated meanings (homonyms) was not reliably 2 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/2 doi: https://doi.org/10.25810/s1d0-gj21 polysemy in the mental lexicon 3 different from that from related meanings. klein and murphy conclude that every sense of a word, related or unrelated, must have a separate mental representation, with no shared core meaning between them. pylkkänen et al. (2006) used a similar semantic judgment task for homonyms, polysemes, and semantically related separate words. while subjects performed the task, the authors used the neuroimaging technique magnetoencephalography (meg) to look for any differences in brain activation between the homonym condition and the polyseme condition. the response time results were similar to those in the klein and murphy (2001) study. however, meg measurements taken during the task found brain activation differences between homonyms and polysemes in both the left and right hemispheres, especially as compared to semantically related separate words. in the left hemisphere, polysemes elicited shorter m350 latencies, which has been hypothesized to index lexical activation. conversely, simultaneous activity in the right hemisphere peaked later for senserelated than for unrelated targets, which the authors suggest may be caused by competition between related senses. furthermore, they propose that these patterns of activation could result from separate mental representations for homonyms, but a single representation for a polysemous word, with separate subrepresentations for the related senses. in trying to discover the shape of our mental lexicon, these researchers have focused on comparing homonyms and polysemes. their materials have primarily consisted of nouns, with individual items falling fairly cleanly into one of those categories. however, in order to test those theories that postulate a clear distinction between homonyms and polysemes, one must look at the border between those two categories. verbs are in general more complex than nouns since they perform the task of relating various other items in an utterance. they often have numerous senses. in fact, the more frequent a verb is, the more senses it is likely to have, with some verbs listed in dictionaries with 50+ senses. for this reason, verbs provide an excellent arena for examining the mental representation of word meaning. verbs may provide us with some insight into whether there is a sharp distinction between meanings that are related and those that aren’t. the oxford english dictionary (2002) lists the senses of draw in ‘he drew his gun’ and ‘he drew a picture’ as related, although most people would see these as unrelated. historically, the sense in ‘drew a picture’ developed from the sense in ‘he drew a line in the sand’, itself from ‘he drew the stick through the sand.’ these senses are distantly related to senses of pulling, such as ‘the ox drew the cart,’ which are related to senses of pulling toward, as in ‘he drew the cup toward him’ or extraction, ‘he drew the gun.’ the overlaps in meaning with this verb are quite complex, and it is difficult to say which would have separate mental representations and which would not. in addition, some senses that may have been seen as related by the average person at some point in the past would now be seen as homonyms, with no shared meaning whatsoever. 3 windisch brown: polysemy in the mental lexicon published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 4 similarly, those theories that advocate separate mental representations for each word sense, related or unrelated, must look at the border between one sense and two distinct senses. for example, wordnet (2005) considers the senses of draw in the following sentences as separate. 1. he drew a knife during the fight. 2. he drew a gun during the fight. 3. he drew a card from the pack. 4. he drew water from the well. would all of these have separate mental representations? would the representations of draw in the following two senses be different as well? 5. he drew his colt .45. 6. he drew his smith and wesson .22. many scholars (hanks 2000; kilgariff 1997; krishnamurthy & nicholls 2000) have discussed the difficulty in determining which usages represent the same sense and which different senses. assuming this problem could be solved, a theory of word meaning that assumes different senses are represented separately would predict quite different behavior for recall of same-sense usages (recall of a single mental representation) versus different-sense usages (recall of distinct mental representations). in addition, this theory would predict that accessing closely related senses would be just as difficult as accessing distantly related senses or unrelated senses. if polysemes and homonyms have the same storage and method of processing, degree of meaning relatedness should have no effect on response times or accuracy. if they each have a separate representation, they should behave similarly. understanding how people connect form and meaning is fundamental to understanding language processing and has implications for lexicography, foreign language learning, and computer processing of natural language. in order to discover the nature of these connections, we must look at the full range of meaning relatedness for a single word or form. in this study, we investigated meaning relatedness by looking at the facility with which people move between senses of a word. given the rich meaning nuances of verbs, we focused on this lexical category, choosing verb senses with varying degrees of relatedness, from homonymy through to nearly indistinguishable senses. 4 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/2 doi: https://doi.org/10.25810/s1d0-gj21 polysemy in the mental lexicon 5 2. method 2.1. participants participants were native english speakers over the age of 18. one group of 20 participated in the preparation of materials, and another group of 33 acted as subjects in the judgment task. 2.2. materials four groups of materials were prepared, each consisting of 11 pairs of phrases. each phrase in a pair contained the same verb. the groups comprised (1) homonymy, (2) distantly related senses, (3) closely related senses, and (4) same senses (see table 1 for examples). placement in these groups depended both on the classification of the usages by the lexical resources wordnet and the oxford english dictionary and on the ratings given to pairs of phrases by a group of undergraduates. the raters placed the relatedness of the verb senses in each pair on a scale of 0 to 3, with 0 being completely unrelated and 3 being the same sense. a pair was considered to represent the same sense if the usage of the verb in both phrases was categorized by wordnet as the same and if raters gave the pair an average rating greater than 2.7 (condition avg. 2.87). closely related senses were listed as separate senses by wordnet and received a rating between 1.8 and 2.5 (avg. 2.26). distantly related senses were listed as separate senses by wordnet and received ratings between 0.7 and 1.3 (avg. 0.92). because wordnet makes no distinction between related and unrelated senses, the oxford english dictionary was used to classify homonyms. homonyms were listed as such by the oed and received ratings under 0.3 (avg. 0.26). interestingly, pairs that the oed considered as having related senses were often seen as homonyms by the raters. although this phenomenon pertains directly to the questions addressed here, we eliminated these pairs to maintain the clarity of the categories. as much as possible, arguments of the verbs were chosen to have minimal semantic overlap within pairs. for example, in the same sense condition, broke the window would be paired with broke the plate rather than broke the glass so that the objects were not from the same semantic field. in addition, phrases were placed in past tense in order to avoid the implication of an imperative in some of the phrases (e.g., break the glass!). table 1: stimuli. prime target unrelated banked the plane banked the money distantly related ran the track ran the shop closely related broke the glass broke the radio same sense cleaned the shirt cleaned the cup 5 windisch brown: polysemy in the mental lexicon published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 6 the stimuli from these four conditions were balanced and randomized with pairs containing semantically anomalous, or “nonsense” phrases. these nonsense pairs contained either two nonsense phrases with the same verb (e.g., hugged the juice/hugged the fund ) or one semantically coherent phrase and one semantically anomalous phrase (e.g., coherent then anomalous: scared the fox/scared the dash, or anomalous then coherent: joined the cliff/joined the team). 2.3. procedure we used a one-way repeated measures design in which every subject saw all 11 pairs in each of the four levels of relatedness. the stimuli from the four levels and the anomalous foils were randomized for each subject. subjects viewed the materials on a computer screen connected to buttons for yes and no. subjects were instructed to press the yes button when a phrase was semantically coherent, or “made sense,” and to press no when the phrase was semantically anomalous, or “difficult to make sense of.” they were instructed to answer as quickly but as accurately as possible. a phrase appeared on the screen and remained in view until the subject responded yes or no. after a 300 ms pause, the next phrase appeared. participants responded to all phrases, both prime and target. both response time and accuracy were measured for target phrases. because all target phrases were semantically coherent, the correct response to these phrases was always “yes.” accuracy was measured as the percentage of target phrases receiving a “yes” response. although participants responded to the “nonsense” pairs, response times and accuracy for these phrases were not included in the analysis, as these pairs did not constitute one of the conditions of study. they were included as foils to the test pairs to encourage participants to fully access the meaning of all phrases. 3. results the following results are based on one-way anovas using planned, orthogonal contrasts. response times differed significantly between the groups (f3,30=18.4; p<.0001), as did accuracy (f3,30=50.5; p<.0001). more focused contrasts reveal interesting facilitation or inhibition effects among the groups. when comparing response times between same sense pairs and different sense pairs (all the other conditions, including closely related, distantly related, and unrelated conditions), we found a reliable difference (same sense mean: 1056ms vs. different sense mean: 1272ms; t32 =6.33; p=.0002). we also found better accuracy for same sense pairs (same sense mean: 95.6% correct vs. different sense mean: 78%; t32=7.49; p<.0001). when moving from one phrase to another with the same meaning, subjects were faster and more accurate than when moving to a phrase with a different sense of the verb, whether that sense was related to the first or not. 6 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/2 doi: https://doi.org/10.25810/s1d0-gj21 polysemy in the mental lexicon 7 moving between closely related senses also proved quicker and more accurate than moving between distantly and unrelated senses. closely related sense response times were reliably faster than the mean of distantly related and unrelated pairs (t32=5.85; p<.0005), and accuracy was higher (t32=8.65; p<.0001). a distinction between distantly related pairs and homonyms was found as well. response times for distantly related pairs was faster than for homonyms (distantly related mean: 1253ms, homonym mean: 1406ms; t32=2.38; p<.0001). accuracy was enhanced as well for this group (distantly related mean: 81%, unrelated mean: 62%; t32=5.66; p<.0001). in this case, polysemy did seem to facilitate access to meaning when compared to homonymy. a final planned comparison tested for a linear progression from completely unrelated senses, through distantly related senses, then closely related senses to same senses. although somewhat redundant with the other comparisons, this test did reveal a highly significant linear progression for response time (f1,32=95.8; p<.0001) and especially for accuracy (f1,32=100.1; p<.0001). these results show a smooth progression, with response time fastest for same sense pairs and then increasing at a regular rate as meaning relatedness decreased. accuracy decreased at a regular rate as meaning relatedness decreased. 0 200 400 600 800 1000 1200 1400 1600 same close distant unrelated figure 1: mean response time (ms). 0 20 40 60 80 100 same close distant unrelated figure 2: mean accuracy (% correct). 7 windisch brown: polysemy in the mental lexicon published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 8 4. discussion the results have implications for various theories of lexical storage and processing, which i will address in turn. the theory that all senses have separate representations, with polysemes and homonyms represented in the same way, receives its strongest support from the comparison of same sense pairs to different sense pairs. this theory would predict this to be a fundamental distinction, with same sense pairs strongly facilitated compared to all different sense pairs. the confirmation of this, however, does not eliminate a theory in which related senses share a core representation. such a theory would be entirely compatible with this result. the separate representation theory also predicts that response time and accuracy for closely related pairs should be similar to distantly related pairs and unrelated pairs. if they all have separate representations, access to meaning should proceed in the same way for all. for example, if one is primed with the phrase ‘fixed the radio,’ response time and accuracy should be the same whether the target is ‘fixed the vase’ or ‘fixed the date.’ we found, however, a significant difference between these two groups, with closely related pairs accessed more quickly and accurately than the distantly related and unrelated pairs. a direct comparison of response times for homonyms and distantly related polysemes revealed a facilitory effect for the polysemes. the increased accuracy and faster response times indicate that moving between one sense and a distantly related sense is easier in some way than moving between two unrelated senses. this suggests some sort of difference in representation between polysemes and homonyms, an implication contrary to the separate representation theory. our results are more compatible with shared representation or “single entry” theories, but not conclusively so. these theories postulate a fundamental difference between homonyms and polysemes. unrelated senses (homonyms) are seen as having distinct semantic representations or “entries” in the mental lexicon. related senses (polysemes) are seen as sharing a portion of a semantic representation, with nuances of meaning diverging from that shared portion. this sort of structure is also described as one “entry” for a polysemous word, with “subentries” for different related senses. as stated above, the fact that the meanings of same sense pairs were accessed more quickly and accurately than the meaning of different sense pairs is compatible with this theory. one would expect accessing a single “entry” or full representation to be easier than accessing different “subentries” of a single “entry” or accessing different extensions of a core sense. the strongest support for this theory, however, comes from the significant difference between distantly related senses and homonyms. if homonyms have completely separate representations, one would expect them to be accessed more slowly and less accurately than senses that share a core representation or “entry,” and this is indeed what we found. this finding in particular calls into question the 8 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/2 doi: https://doi.org/10.25810/s1d0-gj21 polysemy in the mental lexicon 9 separate representations theory while supporting a shared representation or “single entry” theory. the discovery of a linear progression among the groups tested suggests a slight alteration or extension of the “single entry” theory. a linear progression could fit with this theory if, in addition to polysemes being housed in different subentries of a single representation, we picture subentries of subentries, or a hierarchy of sense distinctions. in this way, people might access closely related senses more quickly than distantly related senses. they would have a shorter distance to travel in the hierarchy, so to speak, or have less of an already activated semantic representation to overcome. in this case, however, postulating a qualitatively different representation between polysemes and homonyms is less appealing. homonyms show a similar degree of inhibition to distantly related senses as distantly related senses do to closely related, which suggests that the difference between homonyms and polysemes is only a matter of degree, not of kind. in addition, when people rated the phrase pairs for how related the verb senses were, there was a definite lack of consensus on which senses were distantly related and which were completely unrelated. this supports the notion that the distinction between homonyms and polysemes may not be so clear-cut. nevertheless, there may indeed be a fundamental difference between polysemes and homonyms, with different causes for the similar degree of inhibition we saw between groups. the variations in brain activation found by pylkkänen et al. (2006) seem to support the theory that there is a basic difference between polysemes and homonyms. it would be interesting to see if there is further variation in right and left hemisphere activation when polysemes of different degrees of relatedness are tested. the linear progression through meaning relatedness is also compatible with a theory in which the semantic representations of polysemes overlap. rather than polysemes being discrete entities attached to a main “entry”, they could share a general semantic space. various portions of the space could be activated depending on the context in which the word occurs. this structure allows for more coarse-grained or more fine-grained distinctions to be made, depending on the needs of the moment. 5. conclusion the results of this study refine our understanding of the connection between form and meaning. often a single form is used to represent multiple meanings. these meanings can be semantically unrelated or show different degrees of relatedness. several theories have been proposed as to how people store and process the meanings of these words. one theory holds that every sense of a word has a separate semantic representation in the mental lexicon. another influential 9 windisch brown: polysemy in the mental lexicon published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 10 theory holds that related meanings share a portion of their semantic representation, whereas unrelated meanings have separate representations. these theories have been primarily tested by comparing differences in processing time between noun homonyms and polysemes. we used a semantic judgment task to assess the ease with which people move between senses with four degrees of meaning relatedness. in addition, we used verbs rather than nouns because of the variability of verb meaning when the item is placed in context with different arguments. we found that priming differences between the conditions do not support a theory in which each meaning connected to a form has a separate mental representation. such a theory would predict no difference in semantic processing time when moving from one sense to another that is related or one that is unrelated. however, we found significant differences in processing time and accuracy between processing related meanings and unrelated meanings. even distantly related meanings were processed more quickly and accurately than unrelated meanings. theories that postulate separate representations for homonyms and single but subdivided representations for polysemes were compatible with our findings. in addition, the significant linear progression through meaning relatedness that we found most strongly supports theories in which related meanings share varying portions of their semantic representation, or in which related meanings overlap in semantic space. one can imagine varying portions of shared meaning among different degrees of relatedness. closely related senses could share a large portion of their semantic representations, while distantly related senses would have minimally overlapping representations. the sharing of semantic representations may dwindle until no semantic overlap remains, as in the case of homonyms. as we saw with the verb draw, there may be a sequence of meaning relations that one can trace from “they drew close to the fire” to “they drew water from the well” without the connection being obvious when looking at just those two utterances. this sort of structure is compatible with cognitive linguistics theories of family resemblances and fuzzy boundaries in word meaning and concepts (lakoff 1987; rosch 1975). these theories reject the notion that a category can be defined with necessary and sufficient conditions. instead, members of a category (and polysemes can be seen as members of a category) can be more or less prototypical of the category. also, the boundaries between categories can be fuzzy or indeterminate. several analyses have been done of polysemous words that demonstrate these qualities of overlap, extension, and fuzzy boundaries among the words’ senses (brugman 1981). a structure in which the semantic representations overlap allows for the apparently smooth progression from same sense usages to more and more distantly related usages of a word. it also provides a simple explanation for semantically underdetermined usages of a word, a pervasive phenomenon we have not yet addressed. although separate senses of a word can be identified in different contexts, in some contexts, both senses (or a vague meaning 10 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/2 doi: https://doi.org/10.25810/s1d0-gj21 polysemy in the mental lexicon 11 indeterminate between the two) seem to be represented by the same word. for example, “newspaper” can refer to a physical object: “he tore the newspaper in half,” or to the content of a publication: “the newspaper made me mad today, suggesting that our committee is corrupt.” the sentence “i really like this newspaper” makes no commitment either way. one can equally well imagine the following sentence to be “it has nice large print,” or “it always covers local events.” linguists have attempted to discriminate varying degrees of ambiguity in lexical items (cruse 1986) and to develop criteria for determining when ambiguity indicates either simple vagueness or different senses. geeraerts (1993) revealed the inconsistency and unreliability of such tests, suggesting that a sharp distinction between vagueness and distinct senses may not exist. a theory of semantic representations that allows for overlapping representations or shared core representations helps explain this phenomenon. when encountering a word, one can simply access the core representation or activate the center of the semantic space, and only access further nuances if it is necessary. references azuma, t., and g. van orden. 1997. “why safe is better than fast: the relatedness of a word’s meaning affects lexical decision times.” journal of memory and language 36: 484-504. beretta, a., r. fiorentino, and d. poeppel. 2005. “the effects of homonymy and polysemy on lexical access: an meg study.” cognitive brain research 24: 57-65. brugman, c. 1981. “story of over.” m.a. thesis, university of california, berkeley. cruse, d. a. 1986. lexical semantics. cambridge: cambridge university press. geeraerts, d. 1993. “vagueness’s puzzles, polysemy’s vagaries.” cognitive linguistics 4: 223-272. hanks, p. 2000. “do word meanings exist?” computers and the humanities 34: 205-215. kilgariff, a. 1997. “i don’t believe in word senses.” computers and the humanities 31: 91-113. kintsch, w. 2001. “predication.” cognitive science 25: 173-202. ------. 2007. “meaning in context.” in t. k. landauer, d. s. mcnamara, s. dennis, & w. kintsch (eds.), handbook of latent semantic analysis. new york: psychology press. klein, d., and g. murphy. 2001. “the representation of polysemous words.” journal of memory and language 45: 259-282. klepousniotou, e. 2002. “the processing of lexical ambiguity: homonymy and polysemy in the mental lexicon.” brain and language 81: 205-223. 11 windisch brown: polysemy in the mental lexicon published by cu scholar, 2008 colorado research in linguistics, volume 21 (2008) 12 klepousniotou, e., and baum, s. r. 2007. “disambiguating the ambiguity advantage effect in word recognition: an advantage for polysemous but not homonymous words.” journal of neurolinguistics 20: 1-24. krishnamurthy, r. and d. nicholls. 2000. “peeling an onion: a lexicographer’s experience of manual sense-tagging.” computers and the humanities 34: 8597. lakoff, g. 1987. women, fire and dangerous things. chicago: university of chicago press. nunberg, g. 1979. “the non-uniqueness of semantic solutions: polysemy.” linguistics and philosophy 3: 143-184. oxford english dictionary. 2002. the oxford english dictionary on cd-rom. oxford: oxford university press. pustejovsky, j. 1995. the generative lexicon. cambridge: mit press. pylkkänen, l., r. llinás, and g. l. murphy. 2006. “the representation of polysemy: meg evidence.” journal of cognitive neuroscience 18: 97-109. rodd, j., g. gaskell, and w. marslen-wilson. 2002. “making sense of semantic ambiguity: semantic competition in lexical access.” journal of memory and language, 46: 245-266. ------. 2004. “modeling the effects of semantic ambiguity in word recognition.” cognitive science 28: 89-104. rosch, e. 1975. “cognitive representations of semantic categories.” journal of experimental psychology: general 104: 192-233. ruhl, c. 1989. on monosemy: a study in linguistic semantics. albany: state university of new york press. wordnet. 2005. wordnet, version 2.1. princeton, n.j.: princeton university. retrieved from world wide web: http://wordnet.princeton.edu/perl/webwn?s=word-you-want. 12 colorado research in linguistics, vol. 21 [2008] https://scholar.colorado.edu/cril/vol21/iss1/2 doi: https://doi.org/10.25810/s1d0-gj21 colorado research in linguistics 6-2008 polysemy in the mental lexicon susan windisch brown recommended citation microsoft word cril_paper_brown.doc semantic representations of english and german motion verbs semantic representations of english and german motion verbs cover page footnote this work was supported by a student research grant from the institute of cognitive science at the university of colorado, boulder. we thank dr. bhuvana narasimhan, dr. gary mcclelland, jill duffield, les sikos, michael thomas, alison hilger, shaw ketels and david harper for support, advisement and discussion. german data collection was made possible by our contact in germany, brigitte ulbricht. this working paper is available in colorado research in linguistics: https://scholar.colorado.edu/cril/vol23/iss1/3 https://scholar.colorado.edu/cril/vol23/iss1/3?utm_source=scholar.colorado.edu%2fcril%2fvol23%2fiss1%2f3&utm_medium=pdf&utm_campaign=pdfcoverpages colorado research in linguistics. june 2012. vol. 23. boulder: university of colorado. © 2012 by katherine s. phelps and steve duman. semantic representations of english and german motion verbs katherine s. phelps university of colorado at boulder, department of linguistics steve duman university of colorado at boulder, department of linguistics it has been argued that real-world structure constrains the semantic representations of verbs, resulting in cross-linguistic convergence of naming patterns for motion events. this study explores the nature of this real-world structure by manipulating individual features of human locomotion in video stimuli and comparing the responses of english and german speakers in an elicitation task. we show that individual features influence naming patterns and that languages encode these features differently. furthermore, the semantic representations of several german motion verbs sharply contrast with their english equivalents. 1. introduction languages divide the world in different ways. moreover, the boundaries between semantic categories within a particular language are not necessarily fixed. these two factors contribute to a complicated picture in any cross-linguistic comparison of naming patterns. still, such research has yielded strong evidence of convergent naming patterns across languages in domains such as color (berlin & kay 1969; kay et al. 1997), emotion (eckman 1972), body terms (majid, enfield & van staden 2006) and events (majid, boster, & bowerman 2008; malt et al. 2008). there are at least two major factors contributing to this convergence. first, cognitive biases shared by humans may result in similar construals. second, there are salient discontinuities in the world to which humans attend, and it is this real-world structure that constrains naming patterns. this paper will focus on how real-world structure constrains the semantic representation of motion verbs. this work was supported by a student research grant from the institute of cognitive science at the university of colorado, boulder. we thank dr. bhuvana narasimhan, dr. gary mcclelland, jill duffield, les sikos, michael thomas, alison hilger, shaw ketels and david harper for support, advisement and discussion. german data collection was made possible by our contact in germany, brigitte ulbricht. 1 phelps and duman: semantic representations of english and german motion verbs published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 2 malt et al. (2008) show that structure in the world has a strong influence on the naming patterns of motion events. in a cross-linguistic study in which participants were asked to describe human locomotion, the researchers demonstrate that dutch, english, japanese and spanish speakers uniformly mark a biomechanical distinction between ‘walking’ and ‘running’ gaits when naming these events. however, gait is a cluster of co-occurring features and malt et al.’s (2008) data do not indicate which of these features are encoded by motion verbs. also, their study is limited to four languages and should be augmented with data from more languages. the research we report here manipulates cadence independently from other gait features and shows that it is the latter that influence the category boundary between ‘walking’ and ‘running’ terms, but cadence influences naming on either side of the boundary. second, it incorporates naming patterns from a german dialect that suggest malt et al.’s (2008) claim may be too strong. though some german verbs of human locomotion do encode the biomechanical distinction between ‘walking’ and ‘running’, the term ‘laufen’ (translated variously as walk or run) does not. this runs contrary to prior predictions. moreover, the extent to which cadence influences naming may be languagedependent, as some german verbs appear to be less sensitive to manipulations of cadence than their english counterparts. the present study therefore contributes to research of event categorization in two ways. first, it adds to our understanding of the semantic representations underlying motion verbs by providing a concise picture of the gait features speakers attend to when naming human locomotion. second, it compares naming patterns of human locomotion in english and german, revealing unexpected patterns not present in previous cross-linguistic comparisons. 2. naming human locomotion events continuous human locomotion is particularly interesting due to its biomechanical complexity; it is composed of many co-occurring features. these biomechanical features include but are not limited to stride length, knee and elbow bend, foot contact with the ground, and cadence, i.e., the number of steps per unit of time (kiss, kocis & knoll 2004). combined, these features can be described as a person’s gait, or their manner of motion. a speaker may draw on several of these gait features when naming a human locomotion event. importantly, at a particular speed there is a dramatic switch between the clusters of features often categorized as a ‘walking’ gait—a pendulum-type body motion where at least one foot stays in contact with the ground at all times—and a ‘running’ gait—characterized by more elastic, springing movement (alexander 1992), in which both feet are off the ground at once. malt et al. (2008) demonstrate how this real-world structure—namely the dramatic shift in gait—informs the semantic representations of motion verbs in 2 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/3 doi: https://doi.org/10.25810/8awt-1f35 semantic representations of english and german motion verbs 3 dutch, english, japanese, and spanish. while viewing stimuli of a woman on a treadmill at varying speed settings and inclines, participants were asked to fill in the blank in the sentence: “what is the woman doing? she is _____.” the striking finding of this study was the uniformity in responses with regard to the 4.5 to 5.5 mph treadmill settings. for each language, ‘walking’ terms always appeared from 4.5 mph and slower (and never over 4.5 mph) whereas ‘running’ terms always appeared from 5.5 mph and faster (and never under 5.5 mph). as mentioned above, this distinction marks an important gait difference. the authors argue that this crosslinguistic convergence is the result of structure in the world exerting strong influence over naming patterns. this cross-linguistic convergence does not appear to the same extent on either side of the 4.5/5.5 mph boundary. in english, e.g., much more withinlanguage variation for lexemes such as ‘jog’ and ‘run’ was found, where use of the latter increases with an increase in treadmill speed (between 5.5 and 8.5 mph), but it is never used 100% of the time. while the authors acknowledge that there are many features to which speakers may attend, they admit that “the data do not tell us exactly what cues our participants were responding to” (malt et al. 2008, p. 239). through a small manipulation of the video stimuli in study 1and the addition of a german dialect in study 2, we demonstrate the detailed nature of speakers’ semantic representations of gait terms and show how individual features, particularly cadence, can be a driving force behind naming patterns. we suggest that naming on either side of the 4.5/5.5 mph boundary is quite sensitive to cadence. thus, individual features may significantly affect naming patterns in some circumstances (on either side of the 4.5/5.5 mph boundary) and not others (at the boundary marked by other gait elements). 3. study 1: english the first study had two primary goals. the first was to replicate the english findings of malt et al. (2008). for this reason, we created stimuli as similar as possible to their original human locomotion study. the second goal was to explore the nature of the semantic representations that may underlie naming patterns in terms of relevant features encoded by motion verbs. this required a manipulation of the video stimuli so as to manipulate cadence while controlling other gait features. cadence was the manipulation of choice because its manipulation with digital tools was tractable. 3.1. stimuli stimuli consisted of 21 videos of a college student on a treadmill at varying treadmill settings (1 mph increments from 2.5 to 8.5 mph), in three 3 phelps and duman: semantic representations of english and german motion verbs published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 4 different playback conditions. seven of the videos were unmanipulated and shown at normal playback. using final cut express video editing software, the remaining 14 videos were digitally manipulated to be in either ‘slow motion’ or ‘fast motion’. seven videos were manipulated to slow playback, or 20% slower than normal playback. the remaining 7 videos were manipulated to fast playback, or 20% faster than normal playback. 1 the slow and fast playback conditions are the critical manipulation in this study. in the normal playback condition, all features of human locomotion are coordinated. digital manipulation disrupts this coordination by altering cadence, i.e., the number of steps per unit of time, while controlling other gait elements. (we recognize that cadence is a sub-parameter of gait, but for the purposes of this study we refer to cadence as separate from gait, where the latter remains a collection of co-occurring features such as stride length, knee and elbow bend, ground contact, etc.). for example, with the 6.5 mph treadmill setting at normal playback, gait (stride length, knee and elbow bend, ground contact, etc.) and cadence are in sync. in the slow playback, all of these elements except for cadence remain constant. the stride length, knee and elbow bend, and ground contact are all identical to the normal playback condition. however, the cadence is different. there are fewer steps in the same amount of time. in the fast playback condition, there are more stride revolutions than the normal playback condition. in other words, this manipulation allows us to ‘mismatch’ cadence with other gait elements. 3.2. methods stimuli were shown to 30 native english-speaking undergraduates at the university of colorado at boulder. all undergraduates were monolingual in english with limited experience in a foreign language. videos were randomized to prevent order biases from previous videos and were displayed using the online survey system qualtrics. the videos were mixed with 8 distracters featuring the same actor on a treadmill engaging in activities such as crawling and skipping. upon presentation of each video, participants were asked to respond to the following question: “what is the man doing? he is ______.” participants were asked to use as few words as possible when describing the motion, but using more than one word was allowed. moreover, they were instructed to repeat any word 1prior to the study, naturalness ratings for manipulated videos were obtained. nine undergraduates at the university of colorado at boulder were shown video stimuli and asked to answer the question “how natural is this motion?” by providing a rating on a scale from 1 (not natural) to 5 (very natural). the mean naturalness rating for manipulated videos was 2.8, indicating that while the manipulations were noticeable, they were not unnatural. 4 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/3 doi: https://doi.org/10.25810/8awt-1f35 semantic representations of english and german motion verbs 5 they used as many times as they liked. all participants viewed all videos and the data from all participants was included in the final analysis. all responses were grouped based on the head verb. responses such as ‘running’ and ‘running quickly’ were all grouped as ‘running’. the ‘other’ category includes responses that appeared infrequently (five or fewer times) within a treadmill setting, such as ‘meandering’ or ‘moseying’. 3.3. results the results of the first study reveal two important findings. first, the normal playback condition replicates the results of malt et al.’s (2008) study. ‘walking’ terms are used from 2.5 to 4.5 mph and ‘running’ terms are used from 5.5 to 8.5 mph. second, playback condition is shown to influence naming patterns on either side of the 4.5/5.5 mph boundary, such that some videos are named differently depending on playback condition. the overall effect of playback condition was confirmed by a binomial test (p < .005). 2 across all treadmill settings (2.5 to 8.5 mph), playback speed influences naming patterns. for example, at the treadmill setting of 6.5 mph, the term ‘jogging’ is preferred by 83% of the participants for the slow playback condition, 40% for normal playback, and only 5% for fast playback. a comprehensive view of the data can be seen in figure 1. figure 1 provides a very clear picture of the structure that informs these english motion lexemes: cadence seems to be a critical feature in many terms of human locomotion. even when other gait features remain constant, a change in cadence can result in a change of the most common lexeme for that event. indeed, the manipulation causes additional semantic categories to appear such as ‘power walking’ 3 , which is the most common response for 4.5 mph, but only in fast playback. at the very slowest cadence (2.5 mph, slow playback), participants are 2 to conduct the binomial test, 12 undergraduates from cu boulder were asked to provide a speed ranking of the lexemes from study 1 (e.g., ‘running’ was rated as faster than ‘jogging’ by the majority of participants). these speed rankings were then compared to the data in study 1. across treadmill settings, videos in slow playback were more often paired with lexemes that were rated slower than the most common lexemes in normal playback (e.g., ‘jogging’ < ‘running’), while videos in fast playback were more often paired with lexemes that were rated faster than the most common lexemes in normal playback (e.g., ‘running’ < ‘sprinting’). the binomial test compared the number of times the lexeme changed in the predicted direction (e.g. ‘jogging’ < ‘running’ < ‘sprinting’) to the total number times there was a lexeme change due to playback condition. 3 ‘power walking’ was not considered a modified form of ‘walking’, but rather a compound lexical term, in part because ‘power’ in this case does not pattern with other adverbial modifiers of ‘walking’, e.g., ‘quickly’, which can occur preor post-verbally. 5 phelps and duman: semantic representations of english and german motion verbs published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 6 compelled to use a term other than ‘walking’, though there is less agreement as to what that term should be. this accounts for the large ‘other’ category here, consisting of words such as ‘meandering’, ‘sauntering’, and ‘moseying’. figure 1: the english data from study 1. treadmill settings from 2.5 mph to 8.5 mph and playback conditions are slow (s), normal (n), and fast (f). despite their attendance to the playback manipulation, participants did not use the same lexical item to refer to stimuli on either side of the 4.5/5.5 mph boundary. rather, at this boundary, other gait features seem to be more critical than cadence. if cadence alone were a determining feature, we might expect to see ‘running’ terms applied to 4.5 mph in fast playback, but this is not the case. increased cadence in this condition did not ‘override’ the category boundary, nor did decreased cadence in the 5.5 mph slow playback condition. in the english data, this is without exception. study 1 teases apart cadence from other gait elements and shows that a change in cadence influences naming patterns on either side of the biomechanical boundary. while malt et al. (2008) indicate that strong structure in the world influences naming patterns, they are agnostic in terms of which features are attended to. our results indicate that there is clearly ample structure to which speakers can selectively attend, and that a single feature can play a central or peripheral role in driving naming patterns. 6 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/3 doi: https://doi.org/10.25810/8awt-1f35 semantic representations of english and german motion verbs 7 4. study 2: german by manipulating cadence while controlling other gait elements, study 1 provides a concise picture of what external structural elements of human locomotion influence naming patterns. it also opens up the possibility that speakers of other languages will draw upon these structural features differently than english speakers. to explore this possibility, study 2 replicates study 1 in a german dialect. 4.1. stimuli study 2 used the same stimuli as study 1. 4.2. methods the videos were shown to 28 speakers of a bavarian dialect of german known as rieserisch. this dialect is spoken in the ries area, the capital of which is the town of nördlingen. rieserisch is closely related to the more common german dialect of schwäbisch, spoken primarily in the state of badenwürttemberg (schmidt 1898). though there are several important differences between the grammar and lexicon of rieserisch, schwäbisch, and standard german, specific contrasts between semantic representations of human locomotion verbs in these dialects remain largely unexplored. it is possible that study 2’s results could extend to speakers of more standard german, but this hypothesis requires further investigation. speakers ranged in age from 17 to 48. participants viewed the videos on their own computers through the use of the survey system qualtrics. data from 4 speakers were discarded due to incompleteness. of the remaining 24, all but 2 claimed to have relatively good knowledge of english. an additional 13 claimed knowledge of a third language, and 4 claimed knowledge of a fourth. therefore, only 2 or the 24 participants could be described as monolingual. all, however, identified themselves as native speakers of the rieserisch dialect. upon presentation of each video, participants were asked to respond to the following question: “was macht der mann? er _____.” (what is the man doing? he is _____.). again, participants were asked to use as few words as possible when describing the motion, but using more than one was allowed. they were instructed to repeat any word they used as many times as they liked. all participants viewed all videos. again, all responses were grouped based on the head verb. the ‘other’ category includes responses that appeared infrequently (five or fewer times within a treadmill setting), such as ‘spazieren’ (stroll) and ‘bummeln’ (saunter). 7 phelps and duman: semantic representations of english and german motion verbs published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 8 4.3. results to begin, use of the term ‘laufen’ gives rise to three noteworthy observations. first, contrary to malt et al.’s (2008, p. 239) predictions, the term ‘laufen’—translated as either walk or run—was used to refer to stimuli on either side of the 4.5/5.5 mph boundary (see figure 2). close analysis indicates that 6 speakers used ‘laufen’ across the boundary, 14 used ‘laufen’ but did not cross the boundary, and 4 speakers did not use the lexeme at all. therefore, usage of ‘laufen’ across the 4.5/5.5 mph boundary does not seem to be idiosyncratic or limited to one speaker. second, the use of ‘laufen’ seems to be most common at 5.5 mph. figure 2: comparison of english ‘walking’ and ‘running’ with german ‘laufen’ in normal playback. third, ‘laufen’ is used at every treadmill setting. while never the most frequent term in any given condition, ‘laufen’ is used with high frequency overall, equal to that of terms such as ‘gehen’ and ‘joggen’. hypotheses concerning these observations will be addressed in the general discussion. the effect of playback condition for verbs other than ‘laufen’ was confirmed by a binomial test (p < .05), indicating that a change in cadence affected the choice of lexeme. 4 the verb ‘gehen’ (translated as go or walk) seems to behave differently than english ‘walk’. at 2.5 mph, the change in cadence did 4 lexical rankings for the binomial test were provided by 3 additional native rieserisch speakers who did not contribute to the elicitation task in study 2. 8 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/3 doi: https://doi.org/10.25810/8awt-1f35 semantic representations of english and german motion verbs 9 little to change naming patterns (as seen in figure 3). the same is true for 4.5 mph: where english speakers designate a category of ‘power walking’ for 4.5 mph in the fast playback condition, german speakers do not seem to agree on a motion lexeme in this same condition. this suggests that ‘gehen’ may not encode cadence in the same way ‘walk’ does, and therefore its underlying representation may be qualitatively different. figure 3: the german data from study 2. treadmill settings from 2.5 mph to 8. 5mph and payback conditions are slow (s), normal (n), and fast (f). cadence effects for other verbs were similar to their english counterparts. the verbs ‘joggen’ (jog), ‘rennen’ (run), and ‘sprinten’ (sprint) were sensitive to the change in cadence. for example, at 6.5 mph, ‘joggen’ was used by 58% of the participants in the slow playback condition, 42% of participants in normal playback, and 17% in fast playback. with the exception of ‘laufen’, german naming patterns align with those in other languages that mark the biomechanical distinction between ‘walking’ and ‘running’. the term ‘gehen’ only appears at 4.5 mph and slower; terms such as ‘joggen’ and ‘rennen’ only appear from 5.5 to 8.5 mph. as in english, the cadence manipulation did not cause speakers to break this boundary. 9 phelps and duman: semantic representations of english and german motion verbs published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 10 5. general discussion both study 1 and study 2 bring relevant observations to bear on the nature of the real-world structure that informs semantic representations of human locomotion. they also help support and inform the findings of previous studies. we show that cadence is a structural feature to which english and german speakers attend and that this salience is reflected in naming patterns. presumably, underlying concepts of these verbs will also highlight cadence in this way. both studies show the relative roles of cadence and other gait elements in naming patterns of continuous human locomotion. first, previous hypotheses that the biomechanical distinction between a ‘walking’ and ‘running’ gait (the 4.5/5.5 mph boundary) is the primary structure influencing speakers’ lexeme choices are strongly confirmed. the manipulation of playback (and thus cadence) did not cause any speakers to use a ‘walking’ term from 5.5 mph higher or a ‘running’ term from 4.5 mph and lower. clearly, the biomechanical gait distinction is a critical aspect of speakers’ semantic representations of these human locomotion verbs at this point in the continuum of motion. however, results also indicate that english and german lexemes are extremely sensitive to the manipulation of cadence. in other words, digitally manipulating playback causes people to change the lexeme they use to describe the event (with respect to normal playback). this has important implications. first, it demonstrates that, though the 4.5/5.5 mph biomechanical distinction is of great importance, so too are cues of cadence. moreover, the relative weight given to these features when naming depends on where individual events occur in the continuum of motion. at the 4.5/5.5 mph boundary, motion verbs are distinguished by gait features rather than cadence. on either side of this boundary, however, cadence is driving naming patterns to some extent. this is not to say that only cadence drives naming patterns in these regions. rather, it seems that the conjunction between cadence and gait elements (e.g., stride length, knee and elbow bend, ground contact, etc.) is important in the encoding of motion verbs. that is, the semantic space is multidimensional, with cadence (perhaps perceived as speed) set against other gait elements. the combinatory nature of these dimensions gives rise to particular semantic categories, e.g., ‘jog’ is the conjunction between self-propelled, bounce-and-recoil gaits at medium cadence (see malt et al. 2011 for similar treatments of this multidimensional space). furthermore, contextual dimensions are undoubtedly critical (see labov 1973). for example, people may also attend to the weight of the person (under, average, or overweight), where the locomotion is taking place (indoors, outdoors, on a treadmill), or what they are wearing (casual or sports clothes, etc.). therefore, subsequent studies in this domain should take into account the complicated nature of semantic representations. rather than assume a priori that certain structure in the world will determine naming patterns, we suggest that descriptions of semantic representations should proceed by induction, testing how the possible dimensions of semantic space may be encoded in a given language. 10 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/3 doi: https://doi.org/10.25810/8awt-1f35 semantic representations of english and german motion verbs 11 the german verb ‘laufen’, a difficult term for english speakers, is of particular note. it is commonly translated as either walk or run. however, as can be seen in study 2, this translation is not accurate. there is little overlap in terms of the semantic space encoded by english ‘walk’ and ‘run’ in comparison to ‘laufen’. therefore, direct translation is problematic or even impossible. though such a verb is not present in their data, malt et al. (2008, p. 239) suggest that it may be “possible that some [languages] do not have separate words for walking and running gaits.” while german clearly does have words that mark the biomechanical distinction, ‘laufen’ is a frequent verb that crosses the 4.5/5.5 mph boundary. in fact, ‘laufen’ seems to be used most often at these treadmill settings, perhaps indicating a grouping of 4.5 and 5.5 mph to the exclusion of speeds such as 2.5 and 8.5 mph. this is in contrast to previous predictions regarding such a grouping. one possible interpretation of this finding is that ‘laufen’ is a term of general motion. there are at least two reasons why this explanation is not likely. first, use of ‘laufen’ seems to peak at 5.5mph. if it were a term of general motion, then it should be distributed evenly across all treadmill settings. second, ‘laufen’ features similar metaphorical extensions as the english manner verb ‘run’, not of general motion verbs such as ‘go’. this is demonstrated in (1) and (2): (1) die maschine läuft. the machine is running. (2) das wasser läuft. the water is running. these reasons are compelling evidence to dismiss the characterization of ‘laufen’ as a verb of general motion. a second response is that ‘laufen’ is not a manner verb at all. instead, it is aspectual, meaning “put into motion without delay” (cadiot et al. 2006, p. 182). again, the metaphorical extensions above argue against this treatment, as ‘laufen’ is used to denote a continuous state, rather than indicating the placement of the event in time. second, participants only viewed videos in which motion had already begun. the lack of transition from a non-motion state to a motion state in the videos runs counter to this aspectual reading. in other words, the videos do not indicate that the subject was ‘put’ into motion. we propose, instead, that ‘laufen’ is a specific manner verb of continuous human locomotion that simply draws upon different structure than verbs in english, dutch, japanese and spanish. however, due to the lack of response to the cadence manipulation, we are unable to posit which features in particular figure prominently in its semantic representation. these results also have important implications in regard to so-called ‘manner’ and ‘path’ languages (talmy, 1985). with these studies, we show that ‘manner of motion’ is not an unanalyzable primitive, as has been assumed in the 11 phelps and duman: semantic representations of english and german motion verbs published by cu scholar, 2012 colorado research in linguistics, volume 23 (2012) 12 linguistics literature (slobin 1996; talmy 1985) and the psychology literature (gennari et al. 2002; papafragou & selimis, 2010). rather, ‘manner’ has a finegrained structure, and the current studies are an attempt to tease apart this structure (following efforts of e.g., ikegami 1969). 6. conclusion as close follow-ups to previous studies of motion verbs, the two studies presented here have confirmed prior findings while adding to our understanding of the cues to which speakers may attend when naming human locomotion. by manipulating cadence while controlling for other gait elements, studies 1 and 2 demonstrate that cadence is a critical feature driving naming patterns in addition to the biomechanical distinction between ‘walking’ and ‘running’ gaits. the addition of german data provided unexpected results in the form of the lexeme ‘laufen’, which does not follow the trends found in previous literature. this finding highlights the importance of including as many languages as possible when analyzing semantic representations. this study has concentrated on the semantic representations of human locomotion verbs. we believe, however, that these semantic representations may have important implications for underlying conceptual representations, as defined by levinson (1997). it is possible that the differences in how people talk about human locomotion events may also influence how they think about these events. therefore, this work can inform current research (e.g., malt et al. 2011) on the relationship between semantic representations and underlying concepts. references alexander, r. m. 1992. the human machine. new york: columbia university press. berlin, b. and paul kay. 1969. basic color terms: their universality and evolution. berkeley: university of california press. cadiot, p., lebas, f., and visetti, y. 2006. “the semantics of motion verbs: action, space and qualia.” in m. hickman & s. robert (eds.), space in languages: linguistic systems and cognitive categories. amsterdam & philadelphia: john benjamins publishing company. ekman, p. 1972. “universals and cultural differences in facial expressions of emotion.” in j. cole (ed.) nebraska symposium on motivation, vol. 19. lincoln: university of nebraska press. gennari, s., s. a. sloman, b. c. malt, and w. t. fitch. 2002. “motion events in language and cognition.” cognition 83: 49-79. ikegami, y. 1969. the semiological structure of the english verbs of motion. new haven: yale university press. 12 colorado research in linguistics, vol. 23 [2012] https://scholar.colorado.edu/cril/vol23/iss1/3 doi: https://doi.org/10.25810/8awt-1f35 semantic representations of english and german motion verbs 13 kay, p., b. berlin, l. maffi, and w. merrifield. 1997. “color naming across languages.” in c.l. hardin and l. maffi (eds.), color categories in thought and language. cambridge: cambridge university press. kiss, r., l. kocsis, and z. knoll. 2004. “joint kinematics and spatial-temporal parameters of gait measured by an ultrasound-based system.” medical engineering & physics 26: 611-620. labov, w. 1973. “the boundaries of words and their meanings.” in c-j. n. bailey and r. shuy (eds.) new ways of analyzing variation in english. washington, d. c.: georgetown university press. levinson, s. c. 1997. “from outer to inner space: linguistic categories and non linguistic thinking.” in j. nuyts and e. pederson (eds.) language and conceptualization. cambridge: cambridge university press. majid, a. j. boster, and m. bowerman. 2008. “the cross-linguistic categorization of everyday events: a study of cutting and breaking.” cognition 109: 235 250. majid, a., n. j. enfield, and m. van staden, (eds.). 2006. “parts of the body: cross-linguistic categorization.” language sciences, 28(2-3) [special issue]. malt, b. c., e. ameel, s. gennari, m. imai, n. saji, and a. majid. 2011. “do words reveal concepts?” in l. carlson, c. hölscher, and t. shipley (eds.) proceedings of the 33rd annual conference of the cognitive science society: 519-524. austin, tx: cognitive science society. malt, b. c., s. gennari, m. imai, e. ameel, and n. tsuda. 2008. “talking about walking: biomechanics and the language of locomotion.” psychological science 19: 232-241. papafragou, a. and s. selimis. 2010. “event categorization and language: a cross-linguistic study of motion.” language and cognitive processes 25: 224-260. schmidt, f. g. g. 1898. die rieser mundart. munich: j lindauersche buchhandlung. slobin, d. 1996. “two ways of travel: verbs of motion in english and spanish.” in m. shibatani and s. thompson (eds.) grammatical constructions: their forms and meaning. oxford: carendon press. talmy, l. 1985. “lexicalization patterns: semantic structure in lexical forms.” in t. shopen (ed.) language typology and syntactic description: vol. 3. grammatical categories and the lexicon. cambridge: cambridge university press. 13 phelps and duman: semantic representations of english and german motion verbs published by cu scholar, 2012 colorado research in linguistics 6-2012 semantic representations of english and german motion verbs katherine s. phelps steve duman recommended citation semantic representations of english and german motion verbs cover page footnote cril style sheet you came here to get tacos, bro place references as argumentative resources jacob henry university of colorado boulder references to place are made in interactions to serve a variety of needs, including supporting an argument. however, unlike other ontological categories, place is less fixed as a given geographic location can be conceptualized as being many different places at once. this gives rise to competing notions of place and what a place might index. using case studies of language policing from the corpus of language discrimination in interaction (cldi), this paper examines the ways in which place references are invoked in arguments regarding language, belonging, and proper public conduct. place references in interaction seem strongly linked to ideas of institutionality and what behavior or practices are appropriate given the policies of the institution associated with a place. this paper shows evidence supporting a notion that place references can be broadly divided into categories of local and non-local place references, as both types of reference serve to introduce institutionality into an interaction, but they demonstrate different relationships between the speaker and the invoked institution. this paper also shows the ways in which place can serve to index ideas of racial membership and linguistic identity. keywords: language policing, place, reference, conversation, institutionality, conversation analysis 1. introduction reference serves as a powerful tool in interaction, allowing participants to situate themselves in relation to a specific person, time, location, or other ontological category. enfield 2013:433 calls the practice of reference “a matter of selection,” stating that participants have the option of selecting from several varying options when making a reference. thus, we can interpret the practice of making a reference as inherently agentive; if someone is making a reference in a specific way, there must be a reason as to why they have chosen that specific method of making said reference. in interactional studies, much attention has been paid to person references (enfield 2013, raymond 2016, but we can extend this conclusion of agency to other categories, including that of place. work on place reference in conversation has mostly focused on their mechanics: how are they formulated, how they play into the turn design of an interaction, etc. (schegloff 1972, williams 2016, dingemanse et al. 2017). however, there has been much less work done on how place references can covertly index other meanings or ideologies. references to other ontological categories can serve to indicate a stance or status, and i argue that place references similarly can be used to not only signal some ideology, but to effectively argue in support of said ideology. this paper seeks to better understand the use of place reference as an argumentative resource in the context of language policing incidents from the larger corpus of language discrimination in interaction (cldi). in these interactions, both participants (the “policer” and the “policed”) make place references to support their own stances and ideologies regarding appropriate language use. in presenting transcripts of these language policing events, i aim to elucidate how speakers orient to place in both making and undermining an argument. furthermore, i will show how place references can serve to signal other stances on issues related to language policing, such as nationality and ethnicity. 2. place reference in conversation the reasons why a speaker might choose a specific reference form over another are multitudinous (enfield 2013), ranging from indicating specificity (e.g., “some time last week” vs. “last tuesday at noon”) to assigning affiliation (e.g., “martha complained to me” vs. “your sister complained to me”). these examples also make it clear that a speaker can equally choose forms which obscure specific details or refuse to assign affiliation. many of these reference forms draw their specific meanings from the gricean idea of the cooperative principle; if it is assumed that a speaker is adhering to the cooperative principle, their utterances (and therefore their references) are taken as being relevant and well-formed, so we can draw meanings from their specific forms (grice 1975). however, we can see in conversation that certain forms are selected in ways that are not predicted by grice. for example, grice’s maxim of quantity predicts that speakers provide exactly the necessary amount of information in their speech, but speakers often purposefully avoid anaphoric references (thus providing more information than necessary) in conversation to demonstrate agency (raymond et al. forthcoming). what i call “place references” are possible answers to questions of “where” (levinson 2003). these can include references to specific named locations (e.g. country or city names), references to general institutions (e.g., “in this store”), and deictic references like “here” or “there.” schegloff 1972 proposes three broad analytical categories for what he calls “location references”: location analysis, membership analysis, and topic analysis (or activity analysis). location analysis allows speakers to identify or specify an object (“the rock to the left of the tree”), while topic analysis involves connecting a location to a specific activity (“where we ate lunch”). membership analysis arises in contexts where a place reference can be used to indicate something about the social statuses of interactants. however, unlike other ontological categories, place is multi-faceted, and a given geographic location can be conceptualized as being “multiple places” (schegloff 1972). this gives rise to sometimes competing ideas of a place and what that place might represent, meaning that place references serve as a kind of double-edged sword in argumentative interactions. 3. arguments as conversation as a specific form of interaction, arguments involve speakers making, supporting, and disputing claims. in the literature, arguments have been understood as “conflict talk” (grimshaw 1990). often there are competing claims, and speakers must use interactional resources to manage disagreement. arguments are thought to have a minimum of three turns, as the third position is where the original speaker can reject a conflicting stance from the second speaker (dersley 1998, muntigl & turnbull 1998). this third position is also where we begin seeing the original speaker’s face-saving work in response to a face attack from the speaker in second position. an original speaker’s will respond proportionately to the level of opposition from the second speaker: more direct or overt opposition in second position is seen as more damaging, and so the original speaker will respond with a stronger turn three act to support their original position (muntigl & turnbull 1998). we can see this sort of argument sequence in incidents of language policing. in addition to fitting the conventional definition of an argument or conflict, these interactions can easily be described using the claim-counterclaim framework. as we can see in the example below, there is a clear claim-counterclaim structure regarding what language is appropriate to use. at a restaurant, a customer (cha) asserts that the owner (tar) should speak english. (1) cldi 011 (‘i’ve lived in california for 20 years’)excerpt 01 ((overlapping voices)) 02 cha: (i’ve) lived in california for 20 years and you needeenglish 03 is theour first la:nguage. so you need to speak english. 04 tar: i’m sorry if i (didn’t) (.) you know [i’mi’m okay ] 05 cha: [we:ll i’m sorry about] 06 you too.= 07 tar: =i’mwhy? [i mean i’m] 08 cha: [get the fuck] out of my country. 09 ???: hey wha in line 02, the challenger clearly makes a claim regarding the use of english, and in the following turn, the target acquiesces on line 04. in response to this relatively weak opposition, the act by the challenger in line 08 is a weak, albeit somewhat rude, support of the original position. we can also see an overt place reference in the first turn, giving some credence to the idea that place references can serve as an argumentative resource in language policing incidents. 3.1. language policing we can define language policing as one actor regulating (or attempting to regulate) the language of another actor (blommaert et al. 2009). in the previous example, the language policing is overt and between individuals. however, the historical view of some scholars is that language policing occurs through the lens of regimented language policies. it was thought that while governments and governing bodies conduct language policing through specific policies and practices, “an individual’s right to simply to use a particular language in public or private discourse seems uncontroversial” (king 1997: 495). as we can see in all the interactions from the cldi, this is simply not the case. language policing occurs in all spheres of life and is often performed by individuals who are not acting in an official capacity. in fact, more contemporary work on language policing draws a natural connection between the ideologies of large social institutions, like the media, and the language policing actions of individuals (blommaert et al. 2009). the covert and overt policies of these organizations create and enforce normative language ideologies that individual actors can then internalize, leading to the language policing of individuals. the research on how language policing is conducted is largely tied to broad social institutions, such as schools (amir & musk 2013, cushing 2019) and workplaces (hazel 2015). we can also see language policing in specific communities, such as a group of players in the online game world of warcraft (collister 2014). in all of these situations, language policing is centered around ideas of what is or is not appropriate for the space in which an interaction takes place. “appropriateness” is often derived from normative language ideologies that valorize “standard” english (hazel 2015, cushing 2019). however, interactants can also orient to other goals when engaging in language policing, such as trying to enforce a language of study in a foreign language classroom (amir & musk 2013) or creating an inclusive environment by prohibiting certain slurs or derogatory language (collister 2014). in each case, participants are acknowledging and maintaining some sort of connection between their speech and the institution in which an interaction is taking place. speakers are helping to construct the space through their interaction (heritage & clayman 2010). while much language policing research has focused on institutions and specific communities, there does not seem to be much work on the mechanics of language policing in public spaces between strangers. interactants in a classroom or an online community collaboratively construct institutions, but how do interactants adhere to the rules of an institution when they aren’t building said institution? in many examples from the cldi data, we can see a disagreement about what institutions are relevant and what policies those institutions impose. in the following excerpt, we see a passerby (cha) harassing a group of younger people (tar and unk) as they wait at a crosswalk, ignoring their claim to invoke the institution of the “international city.” (2) cldi xx (‘this is an international city’)excerpt 01 cha then don’t speak so loud.= 02 tar no, it’s international ci[ty he:re.] 03 cha [o:h shut ]up. 04 (0.5) 05 tar [you shut up (goat).] 06 unk [o:h shut up you ]fucking racist. 07 cha fuck off (.) i’m in the () and there’s [just so: many ] of you. 08 unk [you are ra:cist] 09 tar oh, (i’m just) a: dinosaur: ok[ay. ] 10 cha [and you] (.) in this example, there is no clear institution that the challenger is orienting to, and thus she does not invoke an institution in supporting her stance regulating the speech of the targets. however, in line 02, one of the targets uses a place reference (“it’s international ci[ty he:re.]”), and in doing so does appear to make a broad appeal to institutional appropriateness. one would imagine that in an international city, many different languages would be spoken, and moreover, this linguistic diversity is part of what makes the city “international.” while we can’t see the full structure of the interaction, it’s clear that the challenger rebukes this notion and refuses to participate in constructing “an international city.” 4. data in my analysis of place reference, i am drawing on video-recorded interactions that are part of the corpus of language discrimination in interaction (raymond et al. in-progress). the cldi is comprised of incidents in which individuals (the “targets”) are harassed in some way for speaking a language other than english in some public space. the harassers (or “challengers”) engage in language policing against a variety of targets in a variety of contexts. these incidents show us real examples of how language policing can be conducted in a public space and in some cases, seemingly devoid of an institutional influence. many of the interactions are videos taken by the target or some bystander, while others come from security camera footage. because of the nature of these interactions, much of the data is made of up by interactions already in-progress. oftentimes, the act of recording serves to escalate the confrontation, and thus we do not see the true first turns of these interactions. however, the resources that the speakers use in making and undermining arguments become evident as the interaction is continued. as argumentative interactions can be shaped by other cultural practices (ikeda 2008), it must be stated that at present, all of the videos in the cldi are from north america. in some cases, the specific location of an interaction is recorded, and the corpus includes interactions from various regions of the united states and canada. 5. analysis place references are made in many of the cldi interactions, and these examples serve to highlight the relationship between place and authority in an argument. i divide the data into two large categories of place references: interactions where a “local” place is invoked (i.e. references about where the interaction is happening) and interactions where a non-local place is invoked. while both kinds of place reference invoke a sense of institutionality, the local vs. non-local dichotomy shows different relationships between the speakers and the institutions they are calling on. 5.1. local place references most of the interactions in the cldi have some reference to a local place, and this is in part due to the relatively flexible nature of place as an ontological category. however, due to this malleability, participants can orient to place in ways that are contradictory. the first case presented below is an example of such a conflict between a challenger and a bystander in a californian taco restaurant. (3) cldi 009 (‘palapas taco rant’)excerpt 01 cas: do you know what you’re gonna get, (says) friday, [(only-) ] 02 cha: [it sa]ys] 03 it in mexican. we’re not in mexico, we’re in america. 04 cas: sir are you[sir why are you yelling.] 05 cha: [this is america, ] notnot spanish. 06 by1: yeah, but you came here to [get ta ]cos bro. 07 cha: [not span-] 08 huh?= 09 by1: =you came to get tacos. 10 (0.8) 11 cha: so what, (1.0) in america, 12 by1: that’s why::, 13 cha: it’s above the border. (0.4) the red, the border. 14 by1: okay but it’s[(if it wasn’t-] (.) if it wasn’t [(for mexico)] 15 cha: [the water line], [↑fuck you, ] 16 by1: [if it wasn’t for mexico] 17 cas: [(>no no no no< listen).] okay, 18 by1: (we’re pretty) [close to each other,] here we can see that the challenger is the first to mention place in line 03 where he says “we’re not in mexico, we’re in america.” the challenger is contesting the use of spanish in this restaurant by situating it in the broader place of america. however, bystander 1 rebukes the importance of america as a place by instead orienting to the local context in line 06 when he says “yeah, but you came here to [get ta ]cos bro.” bystander 1 is making the claim that in the context of the taco restaurant (“here”), spanish is acceptable or even expected. we can see confirmation of this idea in lines 09 and 12: “=you came to get tacos. that’s why::,” interestingly, this impasse between the challenger and bystander 1 is briefly debated in more references to place. the challenger starts to make an argument based on geography and borders in line 13 (“it’s above the border. (0.4) the red, the border.”), while bystander 1 begins to make an argument presumably based on mexico as the origin of tacos in line 14 (“okay but it’s[(if it wasn’t-] (.) if it wasn’t [(for mexico)]”). neither of these arguments get fully fleshed out before the challenger pivots away from place to his personal identity as an american in line 19 and lack of spanish ability in line 22. as we can see in this interaction, both participants realize place can be an effective tool in making an argument or claiming authority, but this point can be undermined by orienting to place on a different level. we can also see these references to place as inherent appeals to institutionality, with the challenger positioning “america” as the institution embodying the language rules relevant to the interaction, and the bystander doing so with “here” to invoke the institution of the taco shop. it is worth pointing out that the challenger is the first to pivot away from the place argument, possibly showing that he recognized the bystander as having a more salient institution, and therefore a stronger argument. the concept of institutionality comes up in other instances of language policing and is sometimes invoked by representatives of said institution. in the following example, which appears to take place in a grocery store, the manager of the store intervenes in the interaction between the challenger and the target, as a separate recorder looks on. (4) cldi 007 (‘then call the police’)excerpt 10 cha: i’m not harassing you.= 11 man: =well well well. 12 tar: okay [i was being harassed-] 13 man: [stop. that language ] is not allowed in this store. 14 man: if-= 15 tar: =she is ha[rassing me, ] 16 man: [you need to go,] 17 (0.2) 18 cha: i’m [ not harassing you::. ] 19 man: [you’re not welcome here if you] do that. 20 tar: there you go.= 21 tar: =you’re telling me [what language to speak.] 22 man: [you are , 23 (0.2) 24 tar: you’re telling [me what language] to speak. 25 rec: [ thank you:, ] 26 cha: yeah.=speak engl[ish. ]= 27 rec: [bye:.]= 28 tar: =don’t [tell me what to do.]= 29 cha: [you’re in america. ]= 30 rec: =bye:. 31 (.) 32 tar: don’t [tell me wha[t to] do.] 33 cha: [ b y e : [ . ] 34 rec: [bye. ] 35 cha: ((looks toward & waves at camera: 2.0 sec)) 36 rec: say hi to eve[rybody. ]= the manager pointedly tells the challenger that her behavior is not acceptable in the institution in line 13, where she says “[stop. that language ] is not allowed in this store.” here, we can again see the reference to place being used in a way to reinforce the institutionality of the interaction. the challenger is in the place of “this store”, where you can face punishment for not following the protocols in place for interactions with other customers. indeed, we see that continuing to break the rules of the place results in the challenger being rejected from the place. in line 22, the manager begins to make the challenger leave the store and says “[you are ,”. this interaction is similar to the previous example in that there are differing perspectives on how place is relevant in the interaction. in contrast to the manager’s focus on the store as the relevant institution, the challenger tells the target to speak english and uses the larger context of america as justification in line 29: “[you’re in america. ]” so, we can see that again these two participants bring up place to support their conflicting appeals to an institution, and thus, their broader arguments: the challenger uses “america” to support her language policing against the target, while the manager uses “this store” to intervene in and stop the language policing. 5.2. non-local place references we can see in the previous examples that references to local places can be taken as appeals to some institution, and by extension, whatever language policy that institution would embody. in this way, place references are clear argumentative resources that interactants make use of in arguments. however, we can also see a few examples where participants bring up some remote place as a way of supporting their argument. in the following example, we can see a challenger make several local and non-local place references as part of her language policing in a restaurant, and the target’s son jon responds to these references directly. (5) cldi 003 (‘yo no estoy ofendiendo a nadie’)excerpt 31 cha: [yeah you goyou go back to yourspgo back to spain] 32 tar: quiet! be quiet! [close your mouth!] i speak english too. 33 cha: [go back to spain ] 34 (0.5) 35 cha: well [i ] [(see) in america-)] 36 tar: [huh?] okay? [i speak english ] 37 (0.9) 38 tar: nat good, i speak english 39 (0.6) 40 tar: i job in this country. i: job in this country. 41 [pucking you] pucking you [idi (.) you idi i job. i: 42 cha: [ahhhhhhhhhh] [see? thisthis is what youve 43 tar: job] 44 cha: got] 45 jon: =yeah its her fault for doin that, 46 (0.2) 47 cha: [we speak english in the united states ] … 68 jon: and shes not from spain by the way 69 cha: thats [wi] [thats where spain is] frspanish is 70 jon: [so::] dont [be racist like that ] 71 cha: from spain 72 (1.2) 73 jon: you cant be doin th[at, thats racism ] 74 cha: [ive been to spain] so i know … 127 cha: [i dont ]want spanish [ive been to spain ma 128 jon: [o↑ka:y she 129 cha: yeyou havent] been to spain and i have, 130 jon: speaks english] 131 (0.3) 132 jon: i have been to spain befo:re 133 cha: an dyou speak spanish there,[dyou go over there ]and speak 134 jon: [and i speak spanish] 135 cha: russian? we can see here that the challenger is assigning an affiliation to the target by telling the target to “go back to spain” in line 33. she also makes references to united states to justify her actions, as we can see in line 47: “[we speak english in the united states ].” these practices function similarly to the previous examples: the challenger is orienting to “the united states” as the relevant institution from which language ideologies should be derived and is associating the target’s use of spanish with the foreign institution of spain. what is more interesting is that the challenger talks about her own familiarity with this non-local place to strengthen her argument. in lines 69 and 71, the challenger makes the claim “thats [wi] [thats where spain is] frspanish is from spain.” she then goes to claim expertise over the relationship between spain and spanish in line 74 with “[ive been to spain] so i know.” the challenger is claiming to have knowledge over the language and practices of spain due to her having been there, and she is making this claim in such a way as to present a sense of authority in the interaction. of note, she makes this claim after being corrected by jon in line 68: “and shes not from spain by the way.” later in the interaction, the challenger returns to this claim and specifically positions herself in contrast with jon as not having her level of experience. in line 129, she says “yeyou havent] been to spain and i have,”. recognizing this as a challenge to his expertise in the interaction, jon responds in line 132 with “i have been to spain befo:re”, allowing him to occupy the same ground as the challenger. the challenger then seems to use jon’s claim about having been to spain as a way to get him to acknowledge that places are tied to language norms, as she responds to his claim in line 133 and 135 by saying “an dyou speak spanish there,[dyou go over there ]and speak russian?” this could also be taken as a challenge to jon’s knowledge of supposed language norms. thus, we can see both participants orient to familiarity with a place as relevant to their stances in the interaction, even if that place is removed from where the interaction is occurring. this would lend support to the idea that a place can index certain language ideologies and by claiming familiarity with a place, a participant can thus claim expertise over what language policies apply to an interaction. this example is similar to the data presented in transcript (1), where the challenger makes a statement about her expertise by saying “i’ve lived in california for 20 years.” we can see that this statement is repeated several times through the interaction as she continues to try to support her argument. (6) cldi 011 (‘i’ve lived in california for 20 years’)excerpt 01 ((overlapping voices)) 02 cha: (i’ve) lived in california for 20 years and you needeenglish 03 is theour first la:nguage. so you need to speak english. 04 tar: i’m sorry if i (didn’t) (.) you know [i’mi’m okay ] 05 cha: [we:ll i’m sorry about] 06 you too.= 07 tar: =i’mwhy? [i mean i’m] 08 cha: [get the fuck] out of my country. 09 ???: hey wha 10 (variety of voices exclaiming, saying wow) 11 by1: wo::w 12 tar: (so do you thinkso you are) us citizen?= 13 =[so you are gettingso what you]= … 37 cha: [but you’re] in ame:rica? and you need to speak e:nglish. 38 tar: thewhatwhat i’m what i’m doing. what do youwhat do you 39 thi:nk i’m doing, 40 by2: if you’re gonna be racist here you’re gonna leave. 41 ???: (yeah what are you) 42 cha: i’m not racist 43 by2: now now that’s racist. this this man takes care of me >you gon 44 get out of here< if you gon talk like that. don’t do that here. 45 tar: [what do you think i’m doing] 46 cha: [i lived in california ] for twenty years 47 ???: (xxx calm down) 48 by2: >shut up.< don’t talkdon’t talk to these people likethese are 49 good people. don’t do that. [you canyou can leave.] 50 cha: [i lived in california ] for twenty 51 years. [i have the kno:wledge ] the challenger makes the same statement about living in california in line 02 and again in lines 46 and 50, where she goes on to make the statement “i have the knowledge,” thus making a direct connection between expertise with a place and some other kind of knowledge. notably, this interaction occurred in west virginia, so this isn’t a claim about being “from here,” or we would imagine her using specifically west virginia, or a broader geographic descriptor (e.g. “i’ve lived in america/this country/here for 20 years.”). the other instance where place is noted as being important in this interaction is in line 37, where the challenger asserts to the target that “[but you’re] in ame:rica? and you need to speak e:nglish.” thus, this seems to be another example of the challenger positioning a place associated with spanish (or spanish-speakers) as a contradictory institution to “america,” which would prescribe a language policy of mandatory english. of note, it seems that these references to place are also doing some facework as the challenger’s argument is undermined by the target and bystanders. we can see in line 40 that bystander 2 is the first to label the challenger as racist: “if you’re gonna be racist here you’re gonna leave.” the challenger immediately rejects this label in line 42 by saying “i’m not racist,” which shows us that she does not want to be called a racist despite her actions. bystander 2 doesn’t accept this claim and still labels the challenger racist in line 43 by saying “now now that’s racist,” so the next thing that the challenger says in line 46 that she “lived in california for 20 years.” she repeats this in line 50. it’s interesting to note that lines 42, 46, and 50 are sequential for the challenger; she doesn’t say anything else between these lines, giving us something like “i’m not racist. i lived in california for twenty years. i lived in california for twenty years. i have the knowledge.” if we apply the pattern identified by montigl and turnbull 1998 where a more face-damaging rebuttal (like accusing someone of being racist) is met with an act that strongly and overtly supports the original claim, this would indicate that a place reference (and the implied knowledge that it signals) is seen to be a strong argumentative resource, at least by the challenger. 6. discussion we can clearly see participants in these language policing incidents use place references to support their arguments. what’s interesting is that there seems to be a distinction in how participants orient to local place references vs. non-local place references. in the interactions regarding local place references, targets and their allies are forced to address the perceived tensions in “violating” the language policy of the place invoked by the challenger. often, this is done by reframing the place reference to focus on a more salient institution, like we see in (3) with the bystander’s reference to the taco shop and (4) with the manager’s reference to the grocery store. moreover, it appears these incidents are racially motivated, as speakers will often invoke racial categories or slurs during the interaction. challengers are almost always white, and frequently mark themselves as “american,” “canadian,” etc. in contrast, challengers often say that targets are “foreigners” who “don’t belong.” taking this into consideration, local place references do some work to eliminate talking about more emotionally charged topics like race and nationality. these appeals to a local institution seem to draw on the idea that no matter who you are, if you inhabit a certain space, you agree to adhere to the policies of that place. on the other hand, non-local place references seem to allow the challenger to draw a sort of dichotomy between the institution local to the interaction and the institution that the challenger perceives the target to be participating in. non-local place references seem to carry less argumentative weight than their local counterparts as they are not immediately rebuked by the targets, and they are much less frequent in the corpus. however, these non-local references do seem to more overtly index the challenger’s ideologies about who speaks a certain language. if local place references work to remove race or ethnicity from the challenger’s argument, a nonlocal place reference necessarily centers it. in examples (5) and (6), we see the challenger bring up their familiarity with some place that they associate with the language spoken by the target. this harkens to ideas of raciolinguistics, where perceptions regarding race and language use deeply entwined (rosa 2016, rosa & flores 2017). we can see these non-local place references as attempts by the challenger to assert knowledge about not only a place, but also the language spoken there and, by extension, the people who speak that language. this practice reflects what ricklefs (2021) found in elementary school classrooms with bilingual children. non-bilingual white children were habituated to being seen as the more knowledgeable and legitimate students, but when that was undermined by the presence of a language they did not understand, the children were inclined to claim expertise over that language, even if it was just through a simple word or two. in these language policing incidents, i argue that non-local place references act as similar but more complicated methods in preserving the legitimacy of the white participant as more knowledgeable. 7. conclusion & further questions ultimately, this analysis serves as a preliminary examination of how place references function in these language policing incidents. it is clear that they carry some weight as argumentative resources, and that the type of place reference (local vs. non-local) will be taken up in different ways by other participants. local place references are best understood as appeals to the intrinsic language policies of some relevant institution, while non-local place references seem to function as a way for the challenger to express expertise about some oppositional institution. however, these cases do not truly answer questions about how institutionality is constructed in truly public spaces (that is, outside of a shop/restaurant/etc.). there is also work to be done regarding specifically where in the sequence of a language policing interactions place references are made, and if the “placing” of a place reference affects its effectiveness as a resource. further examination can also analyze how non-local references function if they are brought into the interaction by the target instead of the challenger. references amir, alia, & nigel musk. 2013. language policing: micro-level language policy-in-process in the foreign language classroom. classroom discourse, 4(2), 151–167. blommaert, jan, helen kelly-holmes, pia lane, sirpa leppanen, mairead moriarty, sari pietikainen, & arja piiranen-marsh. 2009. media, multilingualism and language policing: an introduction. language policy, 8, 203–207. collister, lauren b. 2014. surveillance and community: language policing and empowerment in a world of warcraft guild. surveillance & society, 12(2), 337–348. cushing, ian. 2019. the policy and policing of language in schools. language in society, 49, 425–450. dersley, ian 1998. complaining and arguing in everyday conversation. york, uk: ph.d. dissertation. dingemanse, mark, giovanni rossi, & simeon floyd. 2017. place reference in story beginnings: a cross-linguistic study of narrative and interactional affordances. language in society, 46 (2), 129–158. enfield, n.j. 2013. reference in conversation. in the handbook of conversation analysis, ed. jack sidnell & tanya stivers, 433–454. oxford, uk: blackwell. grice, p. 1975. logic and conversation. syntax and semantics. 3: speech acts, ed. peter cole & jerry morgan, 41–58. new york: academic press. grimshaw, allen d., ed. 1990. conflict talk: sociolinguistic investigations of arguments in conversations. cambridge: cambridge university press. hazel, spencer. 2015. identities at odds: embedded and implicit language policing in the internationalized workplace. language and intercultural communication, 15(1), 141–160. heritage, john, & steven clayman. 2010. talk in action: interactions, identities, and institutions. malden, ma: wiley-blackwell. ikeda, keiko. 2008. a conversation analytic account of the interactional structure of “arguments.” studies in language and culture, 34, 289–304. king, charles. 1997. policing language: linguistic security and the sources of ethnic conflict: a rejoinder. security dialogue, 28(4), 493–496. levinson, stephen c. 2003. space in language and cognition: explorations in cognitive diversity. cambridge: cambridge university press. muntigl, peter, & william turnbull. 1998. conversational structure and facework in arguing. journal of pragmatics, 29, 225–256. raymond, chase wesley. 2016. linguistic reference in the negotiation of identity and action: revisiting the t/v distinction. language, 92(3), 636–670. raymond, chase wesley, rebecca clift, & john heritage. forthcoming. reference without anaphora: on agency through grammar. linguistics: an interdisciplinary journal of the language sciences. raymond, chase wesley, saul albert, cecillia e. ford, barbara a. fox, natalie grothues, jacob henry, olivia h. marresse, megan pielke, & regina gayou tom. in progress. the corpus of language discrimination in interaction. ricklefs, mariana alvayero. 2021. functions of language use and raciolinguistic ideologies in students’ interactions. bilingual research journal. doi: 10.1080/15235882.2021.1897048 rosa, jonathan daniel. 2016. standardization, racialization, languagelessness: raciolinguistic ideologies across communicative contexts. journal of linguistic anthropology, 26(2), 162– 183. rosa, jonathan, & nelson flores. 2017. unsettling race and language: toward a raciolinguistic perspective. language in society, 46, 621–647. schegloff, emanuel a., 1972. notes on a conversational practice: formulating place. studies in social interaction, ed. david sudnow, 75–118. toronto, ontario: the free press. williams, nicholas jay. 2016. place reference and location formulation in kula conversation. boulder, co: ph.d. dissertation. 1. introduction 2. place reference in conversation 3. arguments as conversation 3.1. language policing 4. data 5. analysis 5.1. local place references 5.2. non-local place references 6. discussion 7. conclusion & further questions toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation colorado research in linguistics. june 2009. vol. 22. boulder: university of colorado. © 2009 by [author full names as they appear in the author line]. toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation nicholas williams university of colorado at boulder this paper presents an overview of research on deixis in linguistic anthropology. in line with other recent deixis theorists (e.g. hanks), i suggest that deixis has not yet received sufficient theoretical nor empirical attention. i argue for the centrality of deixis, and demonstrative reference in particular, to an understanding of the fundamentally social and interactional nature of linguistic meaning. as an exercise in the analysis of deixis in interaction, i analyze the use of two nominal demonstratives (ini and itu) in colloquial indonesian conversation. these demonstratives occur in what are known as “placeholder uses,” frequently in the context of a “word search.” several instances of placeholder demonstrative use are analyzed, showing that differing types of “access” (perceptual, cognitive, social) (hanks 2009) to the referent, as well as distinct indexical grounds, are what distinguish the meaning and use of these two demonstratives. these findings point to the importance of interactional data in the analysis of basic linguistic meaning. 1. introduction: indexicality, deixis and a theory of linguistic “meaning” what pervades all work in linguistic anthropology, from boas to silverstein and beyond, is the attempt to articulate the role of language with respect to culture and to understand what linguistic forms and/or practices mean (all problematic terms, as it turns out). what do linguistic forms/practices mean and how is this reflective of or reflected in the culture? this central problem has been approached in a number of ways over the years. early linguists such as boas and greenberg worked from the “language is culture” approach, using recently developed conceptual tools (“phonemes,” “morphemes” and “syntax”) to describe the grammatical structure of “exotic” languages, particularly those native to north america. based on the assumption that “language is culture”, these linguists masqueraded as anthropologists, “claim[ing] to be doing something anthropological by analyzing grammar” (duranti 2003). however, their focus, and the focus of the long tradition of structural and generative linguists to follow, was not so much cultural meaning in the anthropological sense, but rather a far narrower conception of linguistic meaning or semantics as propositional content and truth-values. this “semantico-referential” approach to meaning takes the “symbol” as its prime object of interest. symbols are characterized by their arbitrary association of the “signified” with the “signifier.” symbols or pairings of signifier and signified (later “form-meaning pairs”) are pervasive in language and, in much of linguistics, typically assumed to be the unique characteristic of language that sets it apart from other semiotic systems. 1 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 2 in the groundbreaking 1976 paper “shifters, linguistic categories and cultural description,” michael silverstein set an agenda for linguistic anthropologists by redefining “meaning” in anthropological inquiry. building on the work of peirce and others in semiotics, silverstein provides a set of tools for understanding the role of language in society, with particular reference to how low-level linguistic forms and micro-practices (or “speech events”) are linked to “culture” or larger, macro-level processes of meaning-making. this perspective is presented in stark contrast to both the traditional semantico-referential approach and to the austinian “speech acts” approach. silverstein finds fault with the traditional semantic approach to meaning for its narrow focus on symbols, “traditionally spoken of as the fundamental kind of linguistic entity” (p. 27). if we are to understand language as a vehicle for the expression of cultural meanings, we will have to acknowledge additional sign types other than the symbol, including the icon and the index. while austin’s and related approaches to “speech acts” and language use are helpful in refocusing our attention on the social life of language and away from a narrow view of propositional meaning, it is flawed in that it assumes a basic level of semantic-referential structure. onto this basic layer of meaning, philosophers like austin and searle “tack on” the “performative use” of these basic linguistic categories. according to silverstein, this entirely misses the point that reference is itself a performative act. there is nothing done in language that is not performative. if anything, reference is a relatively marginal type of action performed through the use of language (but with help from other semiotic systems such as gesture). this old critique bears reconsideration in light of much recent work in “interactional linguistics” which discusses the interactional “use” of various linguistic and/or grammatical “resources” to perform “actions” in conversation. we will leave this issue aside for the moment to examine the rest of silverstein’s theory. however, the notion of reference as a performative speech act will be important later on and we will return to it. as an alternative to this traditional, symbol-obsessed, semanticoreferential and proposition-based mode of analysis, silverstein proposes an approach to meaning in language that focuses on the index, “those signs where the occurrence of a sign vehicle token bears a connection of understood spatiotemporal contiguity to the occurrence of the entity signaled” (p. 27), the meaning of which “always involves some aspect of the context in which the sign occurs” (p. 11). as we will see, this understanding of the index allows for a reinterpretation of all kinds of speech events (including referential speech events) as indexical of some element(s) of the “context in which they occur.” for silverstein, this refocusing will ultimately lead us to the “most important aspect of the ‘meaning’ of speech” (p. 12). in addition to broadening our attention from the semantico-referential to the larger more encompassing field of indexical “meaning” in “culture,” what we gain from silverstein’s approach is a far more nuanced understanding of the index itself, that most important tool for understanding links between language and 2 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 3 context. using the cross-cutting dimensions of presupposing  creative and referential  non-referential, four distinct types of indexes are identified and discussed. the four types are summarized in the following chart, from silverstein (1976): of particular interest to silverstein (1976, 2003) and much later scholarship that draws on the notion of indexicality (e.g. ochs 1990, 1992, inoue 2004, bucholtz 2011, many others) are the “non-referential” indexes, especially those non-referential “creative (performative)” ones. ochs’ (1990) article further developed this concept of the non-referential performative index, adding an additional dimension, direct/indirect. sharing some similarity to silverstein’s (2003) indexical order, this notion of direct/indirect indexes helped further our understanding of the link between performed social identities (e.g. gender) and linguistic practices (e.g. japanese sentence-final particles). as ochs’ pointed out, it is not the case that sentence-final particles in japanese directly index being-awoman or her femininity, but rather that such particles index an “affective stance,” which then indexes a type of “female voice” in japanese speech. in this way japanese sentence-final particles indirectly index gender (conceived of as a performed and socially constructed category of identity). this work is foundational to our contemporary understanding of the links between macro-level social categories and micro-level linguistic practices. however, this is just one direction to take silverstein’s ideas. in the linguistic anthropological literature of the “third paradigm” (duranti 2003), less attention has been given to the referential side of silverstein’s diagram, and even less to the referential presupposing top left corner. what we find in this area are issues typically left to linguists, taken up by semanticists and sometimes pragmaticists, or those scholars probably categorizable in duranti’s (2003) second paradigm. the most notable and well known example of “referential presupposing” indexes is “deixis.” variously defined by different authors, “deixis” for silverstein seems to mean “spatio-temporal deixis,” mainly demonstratives and tense. as he describes it, deixis is “maximally presupposing, in that the contextual conditions are required in some appropriate configuration for proper indexical reference,” and “some aspect of the context … is fixed and presupposed” (p. 34). when using a “deictic” expression (e.g. english this or that, 3 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 4 here or there), speakers make reference to cognitively or perceptually “accessible” (hanks 1990) objects (real or abstract). for silverstein, such accessible objects “exist” for both speaker and addressee and are thus “presupposed,” “otherwise the use of the deictic token is inappropriate” (p. 33). since silverstein’s landmark account of “shifters” and indexicality as crucial parts of the theory of cultural meaning making, a fair amount of progress has been made on that neglected upper left corner of the index diagram. in particular, william hanks (1990, 1992, passim) has argued for a more complex understanding of deictic reference that takes into account not only the object of reference (hanks’ “denotatum”), but also the “indexical ground” and the relation between the two. for hanks and others (e.g. agha 1996), use of deixis is not necessarily presupposing, but rather constitutive of the interactional context of the utterance. agha, for instance, argues, “deictic spatialization effects are not the outcome of ‘coding’ relationships between deictic categories and preexisting spatial realities” (p. 679). that is, deictic expressions do not necessarily (en)code presupposed notions of spatial reality, but rather might be considered more “creative” (in the sense of silverstein 1976). as agha puts it, “deictic usage indexically situates spatial representations in relation to contextual variables whose values are only specified during the course of discursive interaction” (my emphasis). that is, what is presupposed in an act of deictic reference is itself only specified within an instance of situated (talk-in-)interaction, typically through “co(n)textual superposition.” in the remainder of this paper i will further explore the category of deixis as it fits into silverstein’s broader theory of indexicality. rather than relegate it to “traditional linguists” who are likely to give it a semantico-referential treatment, i argue that (spatial) deixis is a crucial feature of situated language use in interaction that requires serious consideration by (linguistic) anthropologists. deictics play a crucial role in connecting instances of language use to the immediate and broader context. deictic reference is accomplished through coproduction with other speakers and is accompanied by use of multi-modal signs (gestures, eye gaze), which are often crucial for interpretation. through the use of data from video recorded conversation in indonesian, i will argue that deictic practice works to constitute (rather than simply reflect or encode spatial aspects of) the immediate context of interaction. i will also show that, in line with hanks (1990), what is fundamental to deictic forms is “access” rather than any spatial notion. finally, i will suggest additional pragmatic effects of deictic use in interaction (in indonesian) that might be interpretable with reference to notions of indexicality and stance (du bois 2007). in this way i aim to situate deictic “referential practices” (hanks 1990) firmly within a linguistic anthropological approach to language and meaning. 2. approaches to deixis 4 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 5 deixis has suffered a history of marginalization in the study of language, alternately being handed off from philosophy to semantics, taken up by pragmatics and occasionally linguistic anthropologists. as a phenomenon it has been largely ignored by most generative and descriptive-typological linguists alike, getting only a cursory treatment in the majority of reference grammars. this is not to say that deixis is lacking a significant literature, which it is not. the tradition of studies of deixis dates back to buhler (1934), at least, and has been followed up by lyons (1977, 1982), fillmore (1982), weissenborn and klein (1982), levinson (1983, 1994, passim), and anderson and keenan (1985), among others. these studies represent mostly philosophical as well as semantic and/or pragmatic approaches to deixis and deictic use. even by 1983, as levinson states in his pragmatics textbook, theoretical approaches to deixis were underdeveloped. since the publication of that textbook, and particularly in the past two decades (1990 – 2010), there has been somewhat of a surge in studies of deixis. this new literature has come from a number of fields, in particular linguistic anthropology (hanks 1990, 1992, 2005, 2009, etc., agha 1996), conversational analysis and interactional linguistics (m. goodwin 1990, c. goodwin 1999a, 1999b) and cognitive psycholinguistic approaches (levinson 1996). additionally, more traditional descriptive semantico-referential and typological approaches have produced accounts of deictics, particularly demonstratives (anderson and keenan 1985, himmelmann 1996, diessel 1999, dixon 2003). a recent overview of the this literature is found in sidnell (1998). schegloff’s (1972) paper on “formulating place” should also be mentioned here, as it informed some of the work on deixis in interaction, although it does not address deixis directly. while in the past few years a number of linguists have published studies of demonstratives and/or deixis in particular languages that draw on these interactionally and ethnographically informed approaches (e.g. bickel 1997, enfield 2003a, 2003b, hayashi and yoon 2006), other work is continuing to be produced which takes the old, simply semantic and problematic view of deixis criticized by hanks (1990, 2009) among others. it will be useful now to briefly summarize past approaches and previous understandings of deixis, as well as the anthropological and interactional critique. this will set us up to look closely at some examples of indonesian deictic demonstrative use in interaction. traditional approaches to deixis are mainly descriptive and typological in nature. fillmore and lyons first discussed deixis from a semantico-referential perspective and contributed to our basic understanding of deixis as a phenomenon distinct from context-referring phenomena such as anaphora and indexicality. more recently, the more “traditional” descriptive-typological approach has been interested in cataloguing the cross-linguistic variation in deictic forms and functions, delineating parameters for universal and language-specific features of deictic systems. this work is exemplified in weissenborn and klein (1982), anderson and keenan (1985), himmelmann (1996), diessel (1999) and dixon (2003), among others. this work has increased our understanding of what can be encoded in deictic systems in a diverse set of languages, thus providing a base 5 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 6 from which to further investigate the everyday use of such forms. himmelmann (1996) is a good example of a study of the functions of diverse deictic (demonstrative) systems across a number of languages. however, there is reason to believe that this well-developed typology of deixis, as well as the “functional” approach to the “use” of these “forms,” is problematic. if we recall silverstein’s (1976) warning, it is not the case that certain semantico-referential aspects of language “structure” are already there, to be left to the formal semanticists and descriptive linguists to discuss, and that these “resources” are then put to “use” in interaction. instead, deictic reference is just one type of language use, one type of performative speech act. while it is true that deictic practices involve certain pragmatic effects, it is misunderstanding the phenomenon to separate the basic “meaning” of the forms from their “use in interaction.” enfield (2003a) has shown very clearly that video data of situated talk-in-interaction is crucial for an accurate analysis of the semantically encoded aspects of demonstratives’ meanings in lao. in the case of lao, it turns out that what is “encoded” semantically is not the expected “proximal” vs. “distal” distinction, but rather a more basic notion of not here for the “distal” form, while the “proximal” form has no semantically encoded meaning. agha (1996) also cautions us against relying on seemingly intuitive and straightforward notions of spatial meanings “encoded” in deictic forms, arguing instead that deictics project “spatial schemas” which are only interpreted and achieve “spatialization effects” through usage in which higher-order superpositions contextualize deictic usage. that is, “aspects of context routinely superimpose spatial construals on deictic usage” (p. 644). this perspective calls to mind, and in fact builds directly upon, the work of hanks (1990, 1992, 2009 and passim), which provides us with a new way of understanding deixis. situated in a phenomenological approach to interaction and grammatical practices (including deictic referential practice), and incorporating rich, ethnographic details relating to the broadly defined context of use, hanks (1990) brings together a new approach to the study of deictic reference, one which presents even more fundamental problems with the traditional, descriptivetypological approach. of particular relevance here is the notion of context and the features of that context “encoded” in or relevant to the interpretation of deictic use. in his 2009 article hanks presents the following simplified diagram to summarize his approach to (spatial) deixis: while previous approaches have focused largely on the object of reference, the denotatum, defined as locatable at some relative distance (“proximal,” “medial” 6 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 7 or “distal”) from the speaker (or “deictic center”), hanks explodes the possibilities by explicitly formalizing the three pieces of any occasion of deictic reference. thus, any occasion of deictic reference points to a denotatum or object of reference (the figure), while also specifying the type of “indexical ground” of reference and the relationship between the two. this is reminiscent of du bois’ (2007) “stance triangle,” which is itself characterized by three reference points, the “stance object” and two separate but dialogically related subjects, connected in one act of stance-taking by relations of alignment, evaluation and positioning. the similarities between these conceptual frameworks will be left here, but a comparison and potential unification of these frameworks deserves much attention. what is crucial to understand about hanks’ approach to deictic reference is the increased number of variables and features relevant for an interpretation of deictic use. hanks (2009) is strongly critical of the “spatialist, egocentric” bias in studies of deixis. rather than understanding deixis as defined in relation to the here-now-speaker deictic center, including the reference to an indexical ground as a relevant parameter allows for different types of deictic reference, including not only “speaker oriented,” but other more “socio-centric” orientations such as “addressee oriented” or “speaker and addressee oriented.” relatively underexplored is the cross-linguistic variation in categorization of the object of deictic reference as, for example, in terms of animacy, gender, number, etc. finally, and most important for the indonesian data we are about to see, the relation between the indexical ground/origo and the object of reference “may be spatial, distinguishing for instance relative proximity, inclusion or orientation. but space is just one sphere of context. other spheres attested in deictic systems include time, perception, memory versus anticipation,” etc. what is fundamental in interpreting instances of deictic reference is the type of “access” indexed by the form. this “access” may be cognitive, perceptual or socially defined. with this background in mind, let us shift to an analysis of some actual examples of deictic reference. the data to be considered come from video recordings of three speakers of indonesian living in boulder, co. the speakers are all bilingual in english and indonesian. 3. demonstrative use in interaction a recent addition to the literature on deixis in interaction is hayashi and yoon’s (2006, reprinted as hayashi and yoon 2010) work on demonstratives and “word formulation trouble” cross linguistically. these authors show that one overlooked and under-appreciated interactional use of deictics (here demonstratives, specifically) is as placeholders in the course of the interactional activity known as “word search.” they delineate three types of such use: placeholder demonstratives (which fill a syntactic slot and maintain the “need for progressivity” and allow speakers to complete their turn without completing the act of reference), avoidance use (placeholders used to avoid producing an 7 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 8 impolite or unacceptable referent) and interjective hesitators (like english “uhm” and spanish “este”, used to indicate occurrence of a word search but not filling any syntactic or semantic role in the utterance). hayashi and yoon show that placeholder demonstratives are pervasive across a range of languages (chinese, korean, japanese, indonesian, finnish, maliseet-passamaquoddy, among others). the prevalence of demonstratives, in particular, used in this type of interactional trouble is attributed to their “pointing function and the invocation of participant access” (hayashi and yoon 2006). their account draws heavily on hanks’s (1990, 1992) analysis of deictic usage, which posits “access” (perceptual, cognitive, social) as the fundamental feature of demonstratives and (spatial) deixis more generally. they show that different demonstrative forms index different types of participant access to knowledge of the intended referent. in korean, for example, medial and distal forms are both used as placeholders during word search. however, while the medial demonstrative indexes “shared access” (indexical ground = speaker + addressee), the distal demonstrative indexes “remote access for speaker” (indexical ground = speaker [only]). these differences in “participant access” follow directly to the hanks-ian notions of relational type and origo or indexical ground, and, in turn, have important consequences for the nature of the following interaction. as hayashi and yoon show, the “shared access” medial forms lead to more overt other-participant involvement in the word search (through offering a possible completion and/or maintaining sustained attention through eye-gaze or other bodily and linguistic practices), while the “remote (from speaker) access” distal forms result in less other-participant involvement and minimal “uh-huh” types of responses. while hayashi and yoon have drawn our attention to the crucial role played by demonstratives and deictic expressions in talk-in-interaction, one potential drawback of this approach is that we are still starting with form, some kind of basic underlying structural distinctions, and then investigating the “use” of these forms in interaction to discover something about “language use in interaction.” as many have shown and enfield (2003a) has so articulately demonstrated, the home of language is face-to-face interaction and it is in interactional data that we will discover the basic encoded semantic distinctions, not only how those distinctions are “put to use” as “interactional resources.” with this in mind, let us turn to some examples of indonesian demonstrative use in interaction. 4. indonesian placeholder demonstratives like english, lao and many other languages, indonesian has a two-term adnominal demonstrative system (this and that), traditionally described as “proximal” vs. “distal” (e.g. sneddon 2006). the forms in indonesian are ini and itu. however, this is just one part of a more elaborate system of (spatial) deixis in indonesian which includes a three-way locative demonstrative system (sini, situ, 8 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 9 sana, roughly: ‘here’, ‘there’ (medial?), ‘over there’ (distal?)), two adverbial demonstratives (begini and begitu, roughly: ‘like this’ and ‘like that’), as well as several other miscellaneous forms including (at least) anu (from javanese, roughly: ‘uhm’/‘whatchamacallit’), nih and tuh (“discourse particles,” grammaticalized forms of ini and itu, with elusive meanings and functions). for the time being we will have to limit ourselves to ini and itu, and in particular, their use as placeholders in the context of word searches. it should be noted that the precise characterization of ini and itu semantics, as well as their function in various forms of discourse and conversational interaction, requires further study. previous work has initiated this (himmelmann 1996, sneddon 2006), but these studies have made limited use of conversational interactional data and do not distinguish the conventional semantics of these forms as from their pragmatic force. the following examples, taken from wouk (2005), demonstrate the placeholder use of ini and itu in relatively straightforward contexts. (1) terus mengenai hadiah-hadiah-nya itu, apa then about redup-gift:gen dem what dari e: e itu, e karang taruna nana sendiri from uh uh dem uh karang taruna nana self "then as for the presents, (were they) what from uh uh that, uh your own karang taruna (name of an organization)." in line 2 the distal demonstrative itu serves as a placeholder for the noun karang taruna nana sendiri. in this example several markers of repair occur, including apa and e: (3x). quite frequently instances of placeholder repair are encountered in which there is no other indication of repair or difficulty recalling the word. for example, (2) o: kalo gitu udah ini dong, lancar oh if like:this already dem emph fluent bahasa inggris-nya language english-gen "oh, in that case (he’s) already this, fluent in english" here the “proximal” demonstrative ini is used as a placeholder for the adjective lancar. we see, then, that placeholder demonstratives in indonesian may stand in for nouns and adjectives (or even verbs or parts of words, not shown here), and 9 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 10 that these placeholder “repairs” may be produced fluently (as in example (2)) without any indication of word formulation trouble aside from the use of the demonstrative. an analysis of the differing functions of placeholders in fluent non-word search production is outside the scope of this paper, but ultimately will have to be accounted for. in fact, a broader approach to demonstratives in all occasions of use will likely shed light on both what is semantically encoded and what pragmatic implicatures underlie the diverse observable occasions of use. following below is an example of prototypical placeholder demonstrative use in a word search from the data collected for this study. note here that in line (1) we see a fluently produced speaker-completed placeholder demonstrative. of more interest here is the use of ini in line (3), where the speaker’s involvement in a word search is clearly indicated by the repetition of ini, the initially cut off production of inand the lengthened vowel [i::] on the second production, the micro-pause in line (4) and the self-addressed ‘whatchamacallit’question, apa nama-nya? at the end of line (3). (3) 1 a: ya udah ini aja, tari-saman kita yeah already this just k.o. dance we latiha:n. [.hh practice “yeah ok, ((let's)) just ((do)) this, tari saman, we(('ll)) practice," 2 l: [m'm= mhm "mhm," 3 a: =sama buka in ini:: >apa nama-nya(˚)?< also open thithis what name-gen "and open in-, ini::, what's it called?" { [a directs gaze at v, far left] } 4 (.) 5 v: booth= booth "booth," 6 a: =booth .hh ntar makanannya:: (.3) kita booth later food-nya we { ^[v nods] } 10 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 11 masak rarame-ra↑me juga cook raall.together also bisa ama [yokke can with y. "booth, then we can also cook the food all together, ((along)) with yokke." in this excerpt the participants are discussing what they would like to do at an upcoming cultural event. they have been discussing types of dance. apparently they have settled on the tari saman. in line (1), a indicates that this topic of conversation is settled and that they should move (ya udah might be glossed as "yeah enough already," but without the negative connotation associated with that english phrase). in line (3) a continues with another turn, suggesting another thing that they might do at the event. while a demonstrative like ini alone might not typically indicate a word search, in this instance the nature of a's production of ini indicates that she is engaged in word search. her first attempt is cut short (in-). in the second production of this proximal demonstrative the final vowel is lengthened quite extensively. this type of "sound stretch" is typical of word searches (hayashi 2003, goodwin and goodwin 1986). immediately following this use of the demonstrative, a utters the common phrase apa namanya, somewhat equivalent to english whatchamacallit, though here it is functioning more like an "interjective hesitator," rather than a placeholder. at the same time, a directs her gaze directly at v, indicating that she is inviting assistance from v for the completion of her word search (see figure 21). the micro-pause following her turn serves to open up the floor and allow v to enter into the word search activity in progress. v's suggestion for a completion is agreed upon by a. a's recycling of the previously searched-for referent, booth, is immediately followed by a quick, but perceptible, nod by v directed at a. this series of gestures and vocal practices points to the careful attention that participants pay to the ongoing activity. in a series of alternating turns, a and v manage to co-construct and complete this word search activity. most importantly here is the observation that the speaker, a, maintains directed eye-gaze with the recipient, v, throughout the activity. we can conclude that word search in indonesian at least has the possibility to begin as a multi-participant activity. word search is not necessarily initiated through the "characteristic" diversion of eye-gaze and production of a "thinking face," as suggested by goodwin and goodwin (1986). 1 in the video stills, speaker v is on the far left, l is in the middle, and a is on the far right. 11 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 12 figure 2. a: [sama buka inini:: >apa namanya(˚)?< of crucial importance here is the use of ini and how we are to account for its use in the context of the word search. while it is clear from the discussion so far that this (and probably any) use of a placeholder demonstrative in indonesia is uninterpretable without consideration of accompanying multi-modal practices (particularly eye gaze), the question remains, what is the function of ini and what does its “meaning” contribute to the unfolding interaction? what can this instance of ini used as a placeholder tell us about the indexical ground, the object of reference and the relation between the two that are “encoded” or “schematized” (agha 1996) in the deictic form in general? the occurrence of an associated gesture and directed eye gaze between participants here is crucial. these multi-modal aspects of the interaction superimpose (in the sense of agha 1996) an interpretation of the deictic reference as accessible to both speaker and addressee. however, it is not clear that the 12 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 13 demonstrative form itself includes addressee in its indexical ground. in the next example we will see how ini can be used in the context of a word search, accompanied by different gaze and gesture practices, to index speaker-only access. we therefore conclude that only speaker is included in the indexical ground for ini. as is typical for an adnominal demonstrative and is clear from the nominal element that replaces ini, the denotatum type or object of reference is a ‘thing’ (as opposed to ‘region’, for example, as for the deictic term “here”). the relational type here appears to be “inclusive,” meaning that the speaker has access to knowledge of the intended referent. following hanks and hayashi and yoon, we could represent this schematically as: table 1. form denotatum type relational type indexical ground ini ‘the one’ inclusive speaker(+addressee?) from this characterization of ini as making reference to a thing in an inclusive relationship to the speaker, the implication follows that this demonstrative, ini, indexes immediate access (of the speaker) to knowledge of the intended referent. as seen in the example above, this indication of speaker-access can be adjusted to include speaker and hearer access to knowledge of the referent through accompanying use of gesture and mutually directed eye gaze. example (4) shows another use of ini as a placeholder, in this case used by the speaker to indicate speaker-only access and avoid overt other-participant involvement in the word search. (4) 1 a: oisn't it crazy? dan orang-orang di swiss { english } and red-person in switzerland gitu-gitu ya? (.5) like.that-red dp 2 mereka ↑tuhe::: ini lho, apa nama-nya? e: setuju they dem uh:: this dp what name-gen uh agree 3 (.3) untuk bayar tax lebih mahal. (.5) karena mereka tahu bahwa de[ngan bayar tax they know that with(by) pay tax "oh, isn't it crazy? and people in switzerland ((and places)) like that, you know? they, uh::, ini lho, what's it called? uh:, agree to pay more expensive [higher] taxes, because they know that with paying taxes..." 13 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 14 5 v: [tax-nya balik lagi ke [tax-nya return again to mere[ka: "the tax ((will)) go back to them again." 6 a: [uh'uh mereka dapat servis: gitu [uh-huh 3p.pl get service like.that "uh-huh, uh-huh, they get service, you know." the excerpt in (4) presents an example of word search, again initiated by a, in which eye-gaze is directed away from the co-participants, resulting in a's own eventual production of the searched-for referent. note that in fragment (6), v engages in co-participant completion to help the talk move forward. here no such assistance is invited, and a ends up completing her own word search. the other two participants display understanding that a is involved in a word search through their uninterrupted eye-gaze directed at a. so, while v and l do not participate vocally in the activity, their eye-gaze serves as an acceptable and appropriate response to the ongoing word search. in this case an interruption or attempted completion by one of these co-participants would indicate a lack of attention to a's current activity expressed through both talk and diverted eyegaze. their silence and directed eye-gaze is thus a salient form of participation in the interaction, while simultaneously they acknowledge a’s indication of speakeronly access. in line (2) a begins the vocal component of her word search with the "filler" e:::, immediately followed by the proximal demonstrative ini. however, an examination of the video data indicates that a has already diverted her gaze by the end of the tuh-. prior to this, a's gaze is directed at the two participants (v in particular see figure 3). the (-) at the end of tuh here indicates an abrupt, probably glottal stop, closure, cutting short the production of this form tuh. this abrupt stop, along with a raised intonation, might itself be the first vocal indication of a word search. during the rest of the word search a maintains eyegaze away from the other participants (see figure 4). during the (.3) second pause following setuju, the first word in her completion of the previous word search, a shifts her gaze back to the other two participants. this excerpt clearly demonstrates the role of gaze in the life of a word search. eye-gaze can be used strategically by the speaker to either invite (example 3) or discourage (example 4) co-participant involvment in the completion of the search. it is not the case (contra goodwin and goodwin 1986), that word searches are characteristically defined by diverted eye-gaze. the diversion or maintenance of mutual eye-gaze between speaker and hearers during word search difficulty are resources used by the speaker for the production of different types of interaction. such multi-modal practices work to superimpose particular interpretations of deictic reference. in this case, direction of eye gaze reinforces the indexing of speaker-only access, 14 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 15 while in example (3) directed eye gaze + a pointing gesture worked to include the addressee(s) in the indexical ground and thus index shared-access. figure 3. a:o isn't it crazy? dan orang-orang di swiss gitu-gitu ya? (.5) { english } and person-red in switzerlandlike.that-red dp 15 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 16 figure 4. 2 mereka ↑tuhe::: ini lho, apa nama-nya? e: setuju (.3) 3p.pldem uh this dp what name-gen uh agree 3 untuk bayar tax lebih mahal. (.5) karena mereka tahu bahwa to/for pay tax more expensive because 3p.pl know that 4 de[ngan bayar tax wi[th pay tax with these two examples, we have tried to make three claims: (1) demonstratives in indonesian are not simply reflective of the immediate spatial context of the utterance, but rather they contribute to the work of constituting the context by indexing particular types of participant access to knowledge of the intended referent; (2) the particular type of participant access indexed follows directly from the indexical ground, denotatum type and relation between these 16 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 17 two features encoded by demonstrative; and (3) these characteristics of the schema (agha 1996) of the deictic form are recoverable from the interaction and should not (and cannot) be based only on exophoric, supposedly “basic” situational use of demonstratives to refer to perceptually accessible objects. it is not the case that “basic” spatial meanings map metaphorically onto non-spatial “endophoric” contexts. instead, the “meaning” of deictic reference forms is constructed in multi-modal interaction. a final example of the “distal” demonstrative itu used as a placeholder will help to reinforce these claims. (5) 1 v: mbak yeny mbak ully malah(an) nari (.) miss y. miss u. in.fact dance buat, (.2) for "y ((and)) u actually dance for ..." 2 a: [i:ya yeah "yeah" {a directs gaze to v} 3 l: [ehh [tunggu mbak= eh 2 wait miss 3 "hey! hold on," {a's gaze is directed to l} 4 v: [itu. that "itu" 5 a: =buat [iya spanyol. for yeah spain/spanish "... for, yeah, spanish ((dance))." {a briefly directs eye gaze to v again} 2 ehh is a frequent attention grabber or marker of interruption. that is, ehh is a resource used to "take the floor" during a conversation. 3 mbak and other address terms are used for second person reference. her the reference is to a, to whom l's utterance is addressed. 17 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 18 in this case v experiences trouble formulating the word spanyol, eventually supplied by a in line 5. to indicate her involvement in a word search, v first produces two brief pauses, a micro-silence followed by a slightly longer (.2) second silence in line 1. this word search differs substantially from the types of word searches involving ini that we have examined previously. in the previous cases, both recipients indicated acknowledgement of the word search through different modes of involvement. in example (3) this involved co-participant completion in the production of the searched-for referent. in example (4) this acknowledgement was indicated by the maintenance of eye-gaze on the part of the recipients. co-participant completion was not a relevant or appropriate form of interaction in (4) because of a's eye-gaze diversion. by diverting her eye-gaze away from the recipients, a indicated that she did not desire assistance in the completion of her word search. on the other hand, v's word search in (5) is interrupted by l's turn at line 3. this indicates that l is not attending to v's experience of word formulation trouble. while a does eventually complete the search through a form of co-participant completion in line 5, this is uttered at a much lower volume than the previous discourse. a also has begun shifting her gaze away from v toward l in response to l's abrupt interruption and attempt to take the floor. a’s offering of a candidate for co-participant completion here might follow from some kind of pressure to complete the reference. while placeholders in indonesian are used for “vague reference” and “avoidance use” (e.g. the frequent use of itu-nya (dem-3.poss) to refer to a male sexual organ, the actual term obviously to be avoided in polite speech), this would not seem to work in this case because the speakers are discussing what kind of dance they will perform in an upcoming event. to refer to the type of dance vaguely with a placeholder like itu is dispreferred in this context since explicit reference to the dance-to-be-performed is needed. if we try to recover the three elements of this deictic form from this example, it seems that in contrast to ini, itu is defined by its non-immediate relational type between a thing (denotatum type) and the speaker (indexical ground). we could schematize this as follows: table 2. form denotatum type relational type indexical ground itu ‘the one’ non-immediate speaker in this case the relation of non-immediacy with reference to the speaker indexes “remoteness of access” for the speaker. since the hearer/addressee is not included in the indexical ground (as with ini), the deictic reference might be interpreted as remote or not to the addressee. in terms of involvement in the ongoing word search and attempt at reference, this means that the addressee may or may not directly participate. thus, as we see in example (5), the addressee (a) becomes involved through directed eye gaze and eventually offering a possible 18 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 19 completion (spanyol). however, this leaves the potential for itu-placeholders to be “filled in” by the speaker alone, without any involvement from the addressee(s). clearly more examples are needed to draw firm conclusions regarding the characteristic features of these two adnominal demonstratives. however, this paper has shown that the “meaning” of demonstratives is not necessarily based on exophoric spatial uses, but can instead be shown to follow from their use in situated interaction. future work will have to show whether these conclusions are valid if we consider the wider range of demonstrative uses across situations, activities and types of discourse. we tentatively hypothesize that these findings regarding the meaning of indonesian demonstratives when used as placeholders will extend to exophoric spatial uses as well, thereby undermining the supposed spatial basis of demonstrative meaning and use (cf. hanks passim). 5. toward a (linguistic anthropological) account of deixis in interaction in this paper i have shown that an anthropological approach (broadly conceived) to demonstrative use, and deixis in general, is needed to account for the “meaning” and interpretation of this important part of language. deixis represents a core example of the contextualized and contextualizing nature of linguistic practice and meaning. detailed micro-analysis of demonstrative use in naturally-occurring interaction can shed light on their meaning(s) and lead to insights unavailable based on the analysis of hypothetical examples. the analysis presented here aims to promote the claim that language is socially constituted and an “emergent” product use in real-time interaction. further evidence will come from more detailed analyses of language use and social interaction. in addition to the semantic and pragmatic “meanings” encoded in these demonstrative forms, it seems quite likely that this practice of placeholder use does additional work for participants in interaction. what i would like to suggest here is that use of a placeholder demonstrative as a type of repair shares something in common with the other-initiated repair discussed by besnier (2010) in his book on the production of gossip on nukulaelae atoll. in this work besnier suggests that the use of non-referential forms (like “he is such a …”) in initial reference position, which evokes other-initiated repair (“who?”, i.e. “who is he?”), works to invite co-participant production of gossip, which ultimately removes culpability and blame from the gossip-initiator. similarly, in the case of placeholder demonstrative use, the use of such a vague reference form (akin to a “recognitional” form, cf. himmelmann 1996) in what looks like “initial” position, occasionally evokes other-participant repair and/or involvement in the word search. this allows speakers to invite other-participant involvement in the production of reference, a fundamental feature of language use and everyday talk and interaction. this avoidance of speaker-only production of reference might relate to the relative rights of speakers and addressees to do reference. this might have to do with shared knowledge among the participants about others’ epistemic 19 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 20 rights. for example, in example (3) v is invited to make reference to the booth because she is much more involved in the preparations for the cultural event being discussed than either a or l is. v, then, holds greater epistemic rights to do the reference in this case, and a’s use of a placeholder demonstrative might be interpreted as an attempt to “downgrade” her own epistemic rights inherent to first position in the adjacency pair (heritage and raymond 2005). the immediately preceding analysis represents the results of a pilot study into the meaning and use of demonstratives in indonesian conversation. future research will need to provide a more comprehensive account of demonstrative meaning, including the relationship between the interactional functions discussed here and the apparent “exophoric” functions commonly proposed as the “basic” function of demonstratives. future work will require a greater amount of data as well as a more ethnographic approach. hanks has drawn linguistic anthropology’s attention to deixis and referential practice as a phenomenon of importance to the field. however, we are still lacking descriptions of deixis and referential practice in a wide sample of languages. echoing hanks (2009), this paper makes the call for greater attention to deixis in linguistic anthropology and other approaches to the study of language (use) and social interaction. 20 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 21 references agha, asif. 1996. “schema and superimposition in spatial deixis.” anthropological linguistics 38: 643-682. anderson, stephen r. and keenan, edward l. 1985. “deixis.” in: shopen, timothy (ed.), language typology and syntactic description, vol. 3, 259-308. cambridge: cambridge university press. besnier, niko. 2009. gossip and the everyday production of politics. honolulu: university of hawai’i press. bickel, b., 1997. “spatial operations in deixis, cognition, and culture: where to orient oneself in belhare.” in nuyts, j., pederson, e. (eds.), language and conceptualization. 46–83. cambridge university press, cambridge. bucholtz, mary. 2011. white kids: language, race and styles of youth identity. cambridge: cambridge university press. b hler, karl. 1934. sprachtheorie: die darstellungsfunktion der sprache. jena: fischer. [1990. theory of language: the representational function of language, translated by donald f. goodwin. amsterdam/philadelphia: john benjamins.] diessel, holger. 1999. demonstratives: form, function, and grammaticalization. amsterdam/philadelphia: john benjamins. dixon, r. m. w. 2003. “demonstratives: a cross-linguistic typology.” studies in language 27: 61–112. dubois, john, 2007. “the stance triangle.” in englebretson, r. (ed.), stancetakinig in discourse: subjectivity, evaluation, interaction. benjamins, amsterdam, pp. 139–182. duranti, alessandro. 2003. “language as culture in u.s. anthropology: three paradigms.” current anthropology 44(3):323-348 enfield, nicholas j., 2003a. “the definition of what-d’you-call-it: semantics and pragmatics of ‘recognitional deixis’.” journal of pragmatics 35, 101–117. enfield, nicholas j., 2003b. “demonstratives in space and interaction: data from lao speakers and implications for semantic analysis.” language 79, 82–117. fillmore, charles j. 1982. “toward a descriptive framework for spatial deixis.” in: jarvella, robert j. and klein, wolfgang, speech, place, and action: studies in deixis and related topics, 31-59. chichester: john wiley & sons ltd. goodwin, charles and goodwin, marjorie. 1986 "gesture and coparticipation in the activity of searching for a word.” semiotica 62(1-2):51-75. 21 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics, volume 22 (2009) 22 goodwin, m. 1990. he-said-she-said: talk as social organization among black children. indiana univ. press. goodwin, c. 1999a. “pointing as situated social practice.” ms, ucla. goodwin, c. 1999b. “action and embodiment within situated human interaction.” ms, ucla. hanks, william f., 1990. referential practice: language and lived space among the maya. university of chicago press, chicago. hanks,william f., 1992. “the indexical ground of deictic reference.” in: duranti, a., goodwin, c. (eds.), rethinking context. cambridge university press, cambridge. hanks, william f., 1996. “language form and communicative practices.” in: gumperz, j., levinson, s. (eds.), rethinking linguistic relativity. cambridge university press, cambridge, pp. 232–270. hanks, william f., 2005. “explorations in the deictic field.” current anthropology 46, 191–220. hanks, william f. 2009. “fieldwork on deixis,” in journal of pragmatics. hayashi, makoto. 2003. “language and the body as resources for collaborative action: a study of word searches in japanese conversation.” research on language and social interaction 36: 109-141. hayashi, makoto and yoon, kyung-eun. 2006. "a cross-linguistic exploration of demonstratives in interaction: with particular reference to the context of word-formulation trouble," in studies in language 30:3, 485-540. heritage, john and geoffrey raymond. 2005. “the terms of agreement: indexing authority and subordination in talk-in-interaction.” social psychology quarterly, 68:1, 15-38. himmelmann, nikolaus. 1996. “demonstratives in narrative discourse: a taxonomy of universal uses.” in: fox, barbara (ed.), studies in anaphora, 205–54. amsterdam/philadelphia: john benjamins. inoue, miyako. 2004. “what does language remember?: indexical inversion and the naturalized history of japanese women.” journal of linguistic anthropology 14(1):39–56. levinson, stephen c. 1983. pragmatics. cambridge university press. levinson, stephen c. 1994. deixis. in r. e. asher (ed.) encyclopedia of language and linguistics, 2:853-57. levinson, stephen c. 1996. “frames of reference and molyneux’s question: cross-linguistic evidence.” in p. bloom, m. peterson, l. nadel and m. garrett (eds.) language and space: 109-169. mit. lyons, john. 1977. semantics. vols. 1 & 2. cambridge: cambridge university press. lyons, j. 1982. “deixis and subjectivity: loquor, ergo sum?” in r. j. jarvella & w. klein (eds.) speech, place and action. john wiley & sons. sidnell, jack. 1998. “deixis.” in: verschueren, jef; östman, jan-ola; blommaert; jan, and bulcaen chris (eds.), handbook of pragmatics 1998, 1-28. amsterdam/philadelphia: john benjamins. 22 colorado research in linguistics, vol. 22 [2010] https://scholar.colorado.edu/cril/vol22/iss1/6 doi: https://doi.org/10.25810/0ghq-wk77 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation 23 silverstein, michael. 1976. “shifters, verbal categories and cultural description.” in basso, k., selby, h. (eds.), meaning in anthropology. school of american research, albuquerque, pp. 11–57.silverstein, michael. 2003. silverstein, michael. 2003. “indexical order and the dialectics of sociolinguistic life.” language and communication 23:193-229. sneddon, james. 2006. colloquial jakartan indonesian. canberra: pacific linguistics. ochs, e. 1990. “indexicality and socialization" in cultural psychology: the chicago symposia, ed. by j. stigler, g. herdt, & r. shweder. cambridge: cambridge university press. ochs, elinor, 1992. “indexing gender.” in: duranti, a., goodwin, c. (eds.), rethinking context. cambridge university press, cambridge, pp. 335–358. schegloff, ea. 1972. “notes on a conversational practice: formulating place.” in studies in social interaction, ed. dn sudnow, pp. 75–119. new york: free press. weissenborn, j rgen and lein, wolfgang (eds.). 1982. here and there: cross linguistic studies on deixis and demonstration. amsterdam/philadelphia: john benjamins. wouk, fay. 2003. “the syntax of repair in indonesian.” unpublished manuscript. university of auckland. 23 williams: toward a linguistic anthropological account of deixis in interaction published by cu scholar, 2010 colorado research in linguistics 6-2009 toward a linguistic anthropological account of deixis in interaction: ini and itu in indonesian conversation nicholas williams recommended citation cril style sheet 1. introduction anyone who has taken a language class at some point in their life has likely used the element of music to at least some degree throughout their learning of the target language (tl), so common has it become in language-learning curricula across the globe. as early as the mid-twentieth century, musical language-learning tools have been implemented with the intention of helping with memorization of features specific to the tl, ranging from individual sounds to colloquial phrases (altessia 2022 and engh 2013). often, they are deemed quite effective to this end, which should not come as much of a surprise, considering the broad span of similarities between language and music. indeed, when attempting to define music, ethnomusicologists often speak using linguistic terminology, arguing that like human languages, varieties of music constitute “self-contained systems” that can be both studied in terms of a broader social context as well as parsed into individual articulatory features (nettl 2005:51). there is also published evidence from the field of psychoacoustics that goes to support this connection, one notable example being the theory of absolute spectral tone color (astc). introduced to the sound and voice sciences by voice pedagogist and researcher ian howell, astc demonstrates that humans conceptualize vowels as pitches by asserting that each frequency (i.e. pitch) within the human hearing range has a vowel-like “color”, a color that remains constant throughout any other acoustic modifications that may be applied to it – just as a musical note played across instruments may vary in its timbre while still sounding the same note (e.g. a middle c being played on a string bass versus being played on a flute) (irene & harris 2022). however, while it may be considered common knowledge that spoken language and music production overlap in their utilization of both the physical and perceptual aspects of sound, there appears to be less known as to just how much having an affinity for processing music can help (or hinder) a person’s ability to process language. it is from this region of uncertainty that my ideas for the following study spawned. my primary intention for completing this project was to observe the extent to which individuals can perceive and reproduce english vowels as represented by synthesized vowels – that is, composite pure tone frequencies generated by praat computer software.1 what i soon became more curious to discover is whether there are any similarities, differences, or other patterns evident among individuals’ perceptions when compared to not only the synthesized vowels themselves, but also individuals’ own attempted productions of the synthesized vowels, and whether these patterns may be dependent on the individuals’ level of musical affinity. using my own musical background as a baseline, i predicted that the higher level of musical background one has, the more likely their perception of a synthesized vowel is to align with their own production of the vowel, regardless of whether they perceive it as the vowel it was actually designed to emulate. this is by the rationale that a higher degree of musical skill generally includes a stronger ability to decipher and closely simulate pure tones. by this same token, i also predicted that even if a person with a higher degree of musical affinity can reproduce what they think they are hearing with more accuracy, it will be more difficult for them to accurately identify the sounds as the natural vowels that they are designed to represent. in the latter case, i proposed that a person’s musicality would work against them in the sense that, because of their intensive training in homing in on individual tones within the context of music (i.e. having been conditioned to identify tones as musical pitches rather than specific spoken vowels), musicians may find it more difficult than non-musicians to perceive the synthesized vowels as the spoken vowels that they are intended to resemble. 2. methods to accomplish my goal of observing individuals’ ability to recognize and produce english vowels out of composite pure tones, i delivered a study wherein i first asked a variety of volunteers to listen to three different sounds, each created as an audio simulation of a spoken vowel. afterwards, i asked volunteers to first determine and then attempt to produce the sound that they heard while i recorded their responses. i was able to collect data from a total of twenty-one participants; however, in constructing my analyses, i ended up using only a portion of what was collected due to the suspected impact of experimental error on some of the participants’ results (see discussion and conclusion). initially, i had intended to recruit the participants such that one third of them would have had some degree of advanced musical study (referred to as formal musicians); another third would have had experience playing music, but not any background formally studying it (referred to as informal musicians); and the final third would have had no experience playing music nor any background in musical study (referred to as non-musicians). while i did end up collecting data from an equal amount of formal and non-musicians, i was only able to retrieve data from a slightly lower number of individuals who fell into the informal category (five speakers as opposed to eight each for the formal and non-musician groups) due to time constraints and lack of accessibility to qualified participants. the category that each participant fell under was determined using knowledge of their occupation and/or field of study as well as the nature of the participants’ identified extra-curricular interests – specifically, whether music is included. as far as other speaker characteristics are concerned, ages ranged from nineteen to sixty years with a mean age of about twenty-five years. out of all who participated, about fifty-two percent identified as female. neither age nor gender was considered as a potential affective variable on the results of this study. i also did not account for any specific dialectal differences between speakers while collecting and analyzing data; however, i did ensure that every participant self-identified as a native english speaker. created using praat, the sounds i provided for my study participants to hear were simply constructed combinations of pure tones layered on top of one another. the pure tones spoken of are equivalent to musical pitches of the same frequency (e.g. middle c, e, and g on a piano). depending on the relationships between them when played simultaneously, these tones have the potential to create a variety of composite sounds, ranging from musical chords (e.g. c major and c minor) to spoken vowels (e.g. the vowels present in the english words “beet” and “bat”). like musical chords, vowels are constructed of multiple tones of specific frequencies layered on top of one another, with the most observable frequencies being referred to as formants. which frequencies are labeled as formants depends on a variety of factors, including the shape of the human vocal tract upon the vowel’s production (soundbridge 2019). conversely, it is a vowel’s empirically prescribed formant values that help distinguish it from other phonemes, including other vowels as well as consonants. (see appendix for a comparison between spectrograms of musical chords and spoken english vowels, as well as a comparison between the spectrogram of a vowel with that of a consonant). due to praat’s lack of humanistic qualities in its production of sound, combining the pure tones chosen for this study amounted to highly digitalized-sounding composite tones that were perhaps initially more representative of analog signals than spoken vowels. nevertheless, i chose to create three of these so-called “composite” tones, each which is composed of individual pure-tone frequencies that correlate with the spoken vowel it is intended to emulate. to create each composite tone, i started by generating three “new” sounds in praat, choosing the option of “create sound as pure tone”. upon selecting this option, i was prompted to edit the acoustic features of the tone, including the number of channels used in playback, start and end times, sampling frequency, tone frequency, amplitude, and fade-in/fade-out duration. for the purposes of this study, the only property i chose to manipulate was the tone frequency, which i adjusted to reflect the formant frequencies recorded in prior research for the vowels i chose to emulate (eduhk 20212, piché 1994-97, and soundbridge 2019). out of all the formant values available, i used only the first three formant values listed, as it is these formant values that are known to characterize a vowel the most prominently and consequently help distinguish it from others (eduhk 2021). after generating the three individual tones according to the frequencies from the referenced formant list, i combined the three by creating a new sound file, done by simply highlighting the three tones and selecting praat’s “combine to stereo” option. this ultimately created three digitalized composite sounds meant to emulate naturally spoken vowels, including the high front unrounded vowel /i/, the middle back rounded vowel /o/, and the low front vowel /æ/ (see appendix for spectrograms of the three simulated vowels, as well as a table showing each vowel’s first three formant frequency values). i chose to emulate these particular vowels mainly on the basis of their variation relative to each other in terms of articulation, with each accounting for a different region of the vocal tract (see appendix). i thought that including as wide as an articulatory variety as possible when playing the study sounds for the participants would yield the best results in terms of the participants’ likelihood of perceiving the sounds as differing in quality, thereby presumably increasing their likelihood to assign vowel qualities to them. before listening to the study sounds, each participant was given a document with a list of common english vowels and instructions to indicate which sound they heard for each sound that is played (see appendix). the list of vowels also included an “other” (fill-in-the-blank) option as well as an “unsure” option. each participant was also given additional oral instructions before beginning the activity: after assuring that they had received the aforementioned document, i explained that i would play three different sounds for them, and after each sound, i would ask them to select a sound they thought they heard using the document as a reference. i also let them know that after selecting a sound, they would then be prompted to produce the sound that they heard as i recorded them using the sound recording function of the virtual meeting space zoom.3 a few additional details that i mentioned to almost every participant included the assurance that they would be given unlimited listens for each sound; no choice that they made would be considered “right” or “wrong”; and upon producing the sounds, to not feel compelled to exactly replicate the sound, but rather produce whatever they thought the sound might be trying to emulate, encouraging them to use their normal speaking voice. after all three sounds were played and the participant’s attempts to produce each sound were recorded, the participant was asked to send their document with their written choices for each sound back to me for data input. 3. data and results to analyze speakers’ ability to accurately identify english vowels from the frequency composition of the three synthesized vowels, comparisons were drawn between sounds that participants perceived/produced and the actual sounds that were emulated. cases wherein participants’ perceptions did not match with their own productions of the sounds were addressed as well, leaving room for the consideration of any possible connections between an individual’s ability to accurately identify the target sound and their ability to produce what they perceive. 3.1. identification of sounds synthesized vowel 1: /i/ sound perceived vs. produced total [formal musicians] total [informal musicians] total [all musicians] total [nonmusicians] total [all participants] /i/ perceived 3 3 6 4 10 produced 3 3 6 5* 11 /u/ perceived 2 1 3 --3 produced 2 1 3 --3 /e/ perceived ------2* 2 produced ------1 1 /ʌ/ perceived 1 --1 --1 produced 1 --1 --1 /ɑ/ perceived ------1 1 produced ------1 1 /ɪ/ perceived ------1 1 produced ------1 1 ? produced ------1 1 table 1. study participants’ identification of synthesized vowel /i/. the asterisk * reflects one case wherein a participant’s perception did not match their own production: /e/ (perceived sound) → /i/ (produced sound) as indicated by table 1, most study participants both perceived and produced synthesized vowel 1 as the sound it was designed to resemble, the high front unrounded vowel /i/. the sound with the second highest perception and production cases for this synthesized vowel was the vowel /u/. in these cases, the height of the synthesized vowel was perceived correctly, but backness was not; rather, it was both perceived and produced as a high-back as opposed to a high-front vowel. this could be a result of trying to mimic the sound too closely, resulting in the production of a rounded tone (no rounded high-front vowel exists in english). three participants perceived and produced synthesized vowel 1 in this way, all falling within either the formal or informal musician category. the spectrogram below shows one of these participant’s synthesized vowel 1 production. figure 1. synthesized vowel 1 production: /u/ other than the three participants who perceived and/or produced synthesized vowel 1 as /u/, the predominant response among both the formal and informal musician categories was the vowel it was intended to emulate, with about sixty percent perceiving and/or producing the sound as /i/. in the non-musician group, exactly fifty percent of the participants identified the sound as /i/, with the remaining responses appearing more varied (i.e. including one identification each for the vowels /ɑ/, /e/, and /ɪ/, in addition to one ambiguous production indicated by “?” in table 1). it may be worth considering how despite this slightly higher variation, several participants in the non-musician category still managed to closely match their productions of the synthesized vowel with their own perception of it, even if their perception didn’t quite match what was actually provided. there was only one case wherein a perception of synthesized vowel 1 differed from not only the target vowel that it was designed to represent, but also the participant’s own production of the synthesized vowel (see special case: perception → production discrepancies between all three synth. vowels). this has been attributed to a lack of understanding on the part of the participant for what was expected of them, which may ultimately constitute as an experimental error (see respective section, as well as potential experimental errors). synthesized vowel 2: /o/ sound perceived vs. produced total [formal musicians] total [informal musicians] total [all musicians] total [nonmusicians] total [all participants] /ɪ/ perceived 1 2 3 1* 4 produced 1 2 3 1* 4 /i/ perceived 2 1 3 --3 produced 2 1 3 --3 /ɑ/ perceived 1 --1 1 2 produced 1 --1 1 2 /u/ perceived 1 --1 --1 produced 1 --1 --1 /ʌ/ perceived ------1 1 produced ------1 1 /e/ perceived ------1 1 produced ------1 1 /æ/* perceived 2 --2 1 3 /ɛ/* perceived ------1 1 /ẽ/* produced 1 --1 --1 /ɑ̃/* produced 1 --1 --1 /æ̃/* produced ------1 1 /iʔ/* produced ------1 1 ? produced ------1 1 table 2. study participants’ identification of synthesized vowel /o/. the asterisk * reflects the cases wherein participants’ perceptions did not match their own productions: /æ/ (perceived sound) → /ẽ/, /ɑ̃/, or /iʔ/ (produced sounds); /ɪ/ (perceived sound) → /æ̃/ (produced sound); and /ɛ/ (perceived sound) → /ɪ/ (produced sound) the data representative of synthesized vowel 2 are highly varied, especially when compared with the data recorded for synthesized vowel 1. there is also a higher rate of perception-production discrepancies for synthesized vowel 2, with two of the individuals from the formal musician category producing something different – more nasalized, in both cases – from what they initially perceived and three of the individuals from the non-musician category producing a completely different sound from what they reported as having perceived. while no participant accurately perceived nor produced the sound that synthesized vowel 2 was designed to emulate (the mid-back rounded vowel /o/), a few came relatively close: a formal musician and a non-musician each perceived and produced it as the low back vowel /ɑ/, while another formal musician identified it as the high back rounded vowel /u/, all three correctly interpreting the vowel’s backness, but not the height nor roundness. when trying to determine a predominant sound selected by the participants for synthesized vowel 2, it can be argued that most individuals tended to perceive and/or produce a high front vowel (either /i/ or /ɪ/); however, this assertion might only reasonably apply to the formal/informal musician cohorts, as non-musicians were characterized with a slightly wider range of identified sounds, including the sounds /ʌ/, /ɛ/, and /e/ in addition to the sounds /u/, /ɪ/, /ɑ/, and /æ/, as well as the ambiguous sound “?”. table 3. study participants’ identification of synthesized vowel /æ/. the asterisk * reflects cases wherein participants’ perceptions did not match their own productions: /i/ (perceived sound) → /iʔ/ (produced sound); “unsure” perception → /eɪ/ (produced sound); /ʊ/ (perceived sound) → /u/ (produced sound); /ʌ/ (perceived sound) → /ɑ/ (produced sound); and /i/ (perceived sound) → /ɪŋ/ (produced sound) synthesized vowel 3: /æ/ sound perceived vs. produced total [formal musicians] total [informal musicians] total [all musicians] total [nonmusicians] total [all participants] /i/ perceived 3* --3 2 5 produced 2 --2 1 3 /ɑ/ perceived 1 --1 1 2 produced 1 1* 2 1 3 /u/ perceived ------1 1 produced --1* 1 1 2 /æ/ perceived ------1 1 produced ------1 1 /e/ perceived ------1 1 produced ------1 1 /ɪ/ perceived 1 --1 --1 produced 1 --1 --1 /ʌ/* perceived --1 1 --1 /ʊ/* perceived --1 1 --1 /iʔ/* produced 1 ------1 /eɪ/* produced 2 --2 --2 /ɪŋ/* produced ------1 1 other: “hitting two spoons together” perceived ------1 1 “unsure” perceived 1 --1 --1* ? produced ------1 1 the data collected for synthesized vowel 3 constitute another case of extensive variation among participants in their perceptions and productions of the synthesized vowel, which in this case was the low front vowel /æ/. only one individual both perceived and produced the vowel as such (a non-musician); their production of the vowel as compared to the actual formant structure of the synthesized vowel is shown in figure 2 and figure 3. the first two formant values of each have been included in the captions to reiterate the closeness with which this participant’s production matched the target sound. figure 2. synthesized vowel 3: /æ/ (f1: 689 hz; f2: 1582 hz) figure 3. synthesized vowel 3 production: /æ/ (f1: ~976 hz; f2: ~1650 hz) notice that the greatest marked difference between the synthesized and participantproduced formants is just below 300 hz, a difference that is relatively small compared to all other comparisons drawn. like synthesized vowel 2, synthesized vowel 3 resulted in some perceptions/productions that could be considered more accurate in terms of frontness (as in the case of the common perception of the high front vowel /i/ among participants) as well as height (as in the participants who perceived/produced the low back vowel /ɑ/). there were also two cases of diphthongization in the production of the target vowel. both roughly constituted the diphthong /eɪ/ as in “ate”, and both were produced by individuals in the formal musician category who did not have a clear idea of what they perceived the sound to be. a visual comparison between synthesized vowel 3 and the /eɪ/ sound produced by one of the participants is shown in figure 4 and figure 5. figure 4. synthesized vowel 3: /æ/ figure 5. synthesized vowel 3 production: /eɪ/ (version 1) this participant’s production of synthesized vowel 3 clearly represents the diphthong /eɪ/, as shown by formants 1 and 2: 1 lowers slightly, indicating an increase in height, while 2 rises, indicating a slightly fronter articulatory position. 3.2. perception → production relationships to measure the extent to which participants were able to produce what they thought they heard, comparisons were drawn between the sounds that they indicated to have heard in written form (using the options provided by the study activity document) and the sounds that they actually produced when prompted. each participant’s sound productions were analyzed using the spectrogram feature of praat and were further compared to spectrograms and formant value tables retrieved from external sources to evaluate each recording of this study within the broader context of human vowel production as a whole – that is, the formant patterns empirically measured for each american english vowel (eduhk 2021). note that not all analyses of participant recordings are described here – just those that were deemed the most relevant based on the extent of their discrepancies/similarities with the participants’ perceptions and/or the empirical formant analyses referenced above. perception → production discrepancies for synthesized vowel 2 figure 6. synthesized vowel 2 production: /ẽ/ (slightly nasalized) when attempting to reproduce what they heard for synthesized vowel 2, this participant (of the formal musician cohort) began by describing the sound as "nasalized" before actually producing the sound. a nasalized quality does appear to be present in the production, based on the higher sporadicity of the shown formants. otherwise, it resembles formant values empirically recorded for the vowel /e/ (eduhk 2021) a bit higher in terms of articulation than the /æ/ vowel that was reported as being perceived. figure 7. synthesized vowel 2 production: /ɑ̃/ (slightly nasalized) here is another case wherein a nasalized quality is evident in the production of synthesized vowel 2, indicated by more sporadically positioned formants. while this could possibly just be a result of a following nasal /n/ (this participant produced the target sound within the context of word "lawn"), it is worth noting that two people ended up producing more nasalized vowels for the second sound. for this participant (also part of the formal musician cohort) specifically, it is also interesting that the production of synthesized vowel 2 ended up being a slightly "backer" vowel than what was indicated as having been perceived (the low front vowel /æ/ referred to eduhk 2021 for formant comparison). figure 8. synthesized vowel 2 production: /æ̃/ (nasalized) here is a third case wherein synthesized vowel 2 is produced as nasalized, not to mention significantly lower than what was perceived (formant values of produced sound come closer to /æ/ vowel than /ɪ/ vowel perceived referred to eduhk 2021 for formant comparison). this participant was part of the non-musician cohort. figure 9. synthesized vowel 2 production: /i/ this participant’s production and perception of synthesized vowel 2 differed mainly in terms of articulatory height (production more closely resembled the high front vowel /ɪ/ than the perceived mid-front vowel /ɛ/ referred to eduhk 2021 for formant comparison). participant was also part of the non-musician cohort. perception → production discrepancies for synthesized vowel 3 figure 10. synthesized vowel 3 production: /eɪ/ (version 2) while this participant’s reported perception was unclear (chose “unsure” option), their production seemed to resemble the diphthong /eɪ/. while the formant structures predominantly resemble those typically measured of the vowel /ɪ/ (eduhk 2021), the first formant seems absent at the onset of the vowel (interpreted as about the 110.326 sec. mark), leading to the possibility of an articulatory position that produced a slightly lower vowel (such as /e/). participant was part of the formal musician cohort. figure 11. synthesized vowel 3 productions: /u/ and /ʊ/ while this participant (an informal musician) initially produced synthesized vowel 3 as /u/ (slightly differing from the /ʊ/ sound indicated as being perceived through its higher frontness and lower height), they then repeated it within the context of the word "cook", resulting in the actual production of the vowel /ʊ/. despite this occurrence, the participant seems to have conceptualized both sounds as the same in both perception and production. figure 12. synthesized vowel 3 production: /ɑ/ this participant’s production of synthesized vowel 3 appears slightly backer and lower than perceived (more closely resembles the sound /ɑ/ rather than the /ʌ/ sound perceived, as compared to formant values prescribed by eduhk 2021). it could be argued that this is a result of the participant using their singing rather than speaking voice when producing the sound, leading to the question of whether pitch and/or tone quality can affect perception/production discrepancies. participant was part of the informal musician cohort. special case: perception → production discrepancies between all three synthesized vowels figure 13. synthesized vowel 1 production: /i/ (no glottal stop) figure 14. synthesized vowel 2 production: /iʔ/ (with glottal stop) figure 15. synthesized vowel 3 production: /ɪŋ/ here is an example of a case characterized with extreme differences between what was heard and what was actually produced for all three synthesized vowels, the only case of its kind in this study. the participant (a non-musician) produced synthesized vowels 1 and 2 as significantly higher/fronter than what was perceived (the vowels /e/ and /æ/, respectively), which is suspected to have most likely been a result of misunderstanding/miscommunication of the study document’s sound list and how its options may relate to the sounds played. while this participant’s production of synthesized vowel 3 came closer to what was perceived (being the high front vowel /i/), there was a more explicit nasal feature present that was not included in the reported perception. this nasal quality differs from that apparent in other participants' productions in the way that the vowel formants are mostly steady (rather than sporadic) until the back end of the production, indicating a clearer phonemic distinction (i.e. presence of a fullfledged nasal consonant). perception → production correlation figure 16. synthesized vowel 1 production: /i/ figure 17. synthesized vowel 2 production: /u/ figure 18. synthesized vowel 3 production: /ɪ/ out of all the formal/informal musicians that participated and whose data could be used (i.e. those unimpeded by experimental error), only one produced all three synthesized vowels as closely matching those that they indicated as having perceived, as indicated through the comparison made between the produced formant values with formant values prescribed for the perceived vowels (eduhk 2021). the other three cases wherein all perceptions matched with productions involved participants in the non-musician category. 4. discussion & conclusion 4.1. summary of data/observations within context of hypotheses ultimately, there seem to have been more instances of close correlation between sounds as they were perceived and produced among the non-musicians as opposed to those with some degree of musical background. while this refutes my hypothesis regarding the nature of this relationship, it does make sense considering the extent to which almost every participant in the formal and informal musician categories seemed to have made more of an effort to mimic the sound that they heard, regardless of the vowel that they may have initially perceived. due to the limit that both english and ipa orthography impose on what can be transcribed within the context of natural human speech, it is understandable if any additional details an individual may try to implement when actually producing what they hear differentiate from what they claim to have perceived using the limited letters/symbols they have been provided. what is perhaps even more fascinating to consider is how in many cases, the participants who did appear to add more acoustic features to their sound productions than they described as hearing appeared to have done so unconsciously, perhaps further endorsing their affinity to decode sounds from a perspective centered around musicality as opposed to linguistic meaning. another observation worth mentioning is the fact that any discrepancies in sound production as compared to sound perceptions among participants in either of the two musician categories appear to be a result of such attempts to add to the perceived sound, as opposed to resulting from a complete misinterpretation of the sound, which seemed more likely to occur among the non-musicians. when reviewing the rate at which participants accurately identified the synthesized vowels as the vowels they were designed to emulate, one notable observation that can be made is the fact that out of the three synthesized vowels, synthesized vowel 1 (/i/) ended up with the highest rate of accuracy, as well as having the fewest occasions of perceptions that did not align with productions. this may indicate that there is a direct relationship between the number of accurate identifications of a synthesized vowel and the frequency at which it is produced as it is perceived. because of the extent of variation among participants from all three cohorts in their perceptions and productions of the synthesized vowels, it is difficult to come to a defendable conclusion regarding the effect of one’s musical affinity on their ability to accurately identify the synthesized vowels as the spoken vowels they were designed to emulate. however, when considering the proposed negative relationship between musical affinity and perceptionproduction correlation among individuals, as well as the apparent tendency for individuals to more accurately perceive and produce the target sound when their perception aligns with their production, it does appear more likely that a person with a higher degree of musical background may experience more difficulty in attempting to identify a vowel from a set of raw, pure tone frequencies alone. this supports the theory that i used to formulate my initial hypothesis regarding the effect of musical affinity on sound identification accuracy – that is, due to the extensive audiation training that many formal (and perhaps some informal) musicians have, their ability to simplify composite pure tone frequencies within a context other than music becomes compromised as too often they become stuck on the acoustic properties and therefore oblivious to the linguistic implications. such conclusion may also be supported by previously conducted psychoacoustic research mentioned in the introduction of this paper – in essence, the theory of absolute spectral tonal color. while asserting the notion that individual vowels can be assigned to specific pitches (i.e. frequencies), the theory also illustrates that despite each frequency having a vowel-like color, these vowel-like colors aren't perceived as the same as the composite vowel actually being produced (irene & harris 2022). thus creates an auditory illusion directly involving the mismatching of vowel production and vowel perception, a phenomenon that appears to be central in the results of the present study. 4.2. potential experimental errors study task design: instruction delivery as indicated in the case wherein all three synthesized vowels were perceived and produced with discrepancies, there appears to have been a misand/or lack of communication between me and at least one of the study participants – primarily regarding how/what to produce upon being prompted to reproduce what they thought they heard. as the study progressed and i gained a better understanding of how people would typically respond to the task at hand, i attempted to improve/clarify the instructions – specifically to the effect of ensuring that when trying to produce the sounds, participants were aware that they could produce vowels as they would usually speak (as opposed to mimicking the synthesized sounds exactly as heard). study task design: list of possible sounds it was not until i had already begun facilitating the study task to participants when i realized that i had forgotten to include diphthongs in the list of sounds i provided for participants’ reference on the study handout. similarly, i realized that i could have also included other voiced sounds (sonorants, in particular). i am unsure as to whether this would have affected the overall patterns of perception-production correlations among participants or whether it would have helped them correctly identify the target vowels. condition of audio some participants ended up hearing at least one of the three sounds at a higher amplitude than the original amplitude of 70db due to reported volume issues. i am uncertain as to the exact source of the problem, especially since not everyone experienced the issue; with this in mind, i expect it might have been a result of variation in tech quality across participants. any data collected from a situation involving an increase in amplitude/intensity were excluded from the final analysis. 4.3. study implications and questions for further research i am satisfied to acknowledge that the results of this study – while leaving some conditions of the measured relationships still up for interpretation – ultimately support that there is indeed a relationship between the way an individual processes the sounds of music and how they process the sounds of language. the implications for this assertion are evident not only within the field of linguistics as a whole, but perhaps specifically within the area of language acquisition, as it brings into question how to best treat language learners who may find it especially difficult to learn a language’s sound system due to their stronger inclination towards processing individual speech sounds in terms of perceived “musical” attributes rather than phonemic meaning. the results of this study have also left me with an assortment of further questions that might be worth addressing in the future, including: is there any relationship between an individual’s ability to decipher vowel-like sounds from pure tone collections and their ability to identify musical pitches and/or chords from the same collections (designed like the ones used in this study)? to what extent might other acoustic attributes of the composite pure tones of a sound affect perception/production (e.g. amplitude/intensity, duration, sequence, etc.)? does dialect play a role in a person’s ability to process synthesized speech? what about fluency in other languages (including cases wherein english is not an individual’s first language)? are there similar relationships evident between individuals not only of languages other than english, but also who come from non-western music backgrounds? appendix 1. synthesized, praat-generated vowels used in study task synthesized vowel 1: /i/ synthesized vowel 2: /o/ synthesized vowel 1: /i/ pure tone collection synthesized vowel 2: /o/ pure tone collection synthesized vowel 3: /æ/ pure tone collection formant values used to create synthesized vowels, adapted from eduhk (2021), piché (1994-97), and soundbridge (2019) synthesized vowel f1 (hz) f2 (hz) f3 (hz) /i/ 280 2207 2254 /o/ 408 803 2579 /æ/ 689 1582 1660 2. spectrogram comparisons the below spectrograms illustrate the similarities between sounds prescribed as musical chords (see figure 1) and sounds prescribed as vowels (see figure 2), as well as the difference between vowels and voiced consonants (see figure 3), all from a harmonic frequency perspective. notice the common presence of layered frequencies in the productions of the chords and the vowels, with the dark red marks indicating the most resonant frequencies (i.e. formants). also notice how in the third spectrogram, there is an absence of multiple resonant frequencies during the articulation of the two consonants /b/ and /j/, even though they are still considered voiced phonemes. such emphasizes the importance of vowels’ harmonic properties when distinguishing them from other voiced phonetic sounds. a spectrogram showing six musical chords (root c) played in succession on a piano. chords from left to right are c, c minor, c major-seven, c seven, c minor-seven, c minorseven. https://www.researchgate.net/figure/spectrogram-of-the-piano-chords-set-c-cm-cm7-c7-cm7-cm7_fig2_228525894 part of a spectrogram showing the production of four english words, each with a slightly different vowel at its core. https://soundbridge.io/formants-vowel-sounds/ https://www.researchgate.net/figure/spectrogram-of-the-piano-chords-set-c-cm-cm7-c7-cm7-cm7_fig2_228525894 https://soundbridge.io/formants-vowel-sounds/ a spectrogram showcasing the presence of two different consonants in between the vowel /a/; the top contains the voiced plosive /b/, and the bottom contains the voiced palatal approximant /j/. https://www.researchgate.net/figure/two-examples-of-the-spectrograms-and-electrodograms-used-in-this-studypanels-a-and-c_fig1_10624516 3. american english vowel chart below is a diagram of the monophthongs of american english, organized relative to each vowel’s articulatory position in the human vocal tract. these articulatory positions were considered when determining which vowels to simulate for the study (one vowel from each depicted height as well as one rounded back vowel were included in the effort to represent the most contrast). https://www.researchgate.net/profile/benjamin_bay2/publication/317571199/figure/fig1/as:504926465781760@149739525856 7/for-rhyme-scoring-purposes-we-estimate-vowel-similarity-by-finding-the-distance-between.jpg https://www.researchgate.net/figure/two-examples-of-the-spectrograms-and-electrodograms-used-in-this-study-panels-a-and-c_fig1_10624516 https://www.researchgate.net/figure/two-examples-of-the-spectrograms-and-electrodograms-used-in-this-study-panels-a-and-c_fig1_10624516 https://www.researchgate.net/profile/benjamin_bay2/publication/317571199/figure/fig1/as:504926465781760@1497395258567/for-rhyme-scoring-purposes-we-estimate-vowel-similarity-by-finding-the-distance-between.jpg https://www.researchgate.net/profile/benjamin_bay2/publication/317571199/figure/fig1/as:504926465781760@1497395258567/for-rhyme-scoring-purposes-we-estimate-vowel-similarity-by-finding-the-distance-between.jpg 4. copy of vowel perception study hand-out using the list below, please write what you hear for each sound that is played. (a) /i/ as in ‘beet’ (b) /u/ as in ‘shoot’ (c) /æ/ as in ‘hat’ (d) /o/ as in ‘orchard’ (e) /ɪ/ as in ‘pit’ (f) /e/ as in ‘cake’ (g) /ɛ/ as in ‘bed’ (h) /ʌ/ as in ‘cup’ (i) /ɑ/ as in ‘lawn’ (j) /ʊ/ as in ‘cook’ (k) other: _______ (l) unsure sound 1: __________ sound 2: __________ sound 3: __________ references altissia. 2022. music as an effective tool for learning languages. altissia online language learning platform: https://altissia.org/music-as-an-effective-tool-for-learning-languages/. eduhk. 2021. 2.2 formants of vowels. online corpus-aided english pronunciation teaching and learning system: https://corpus.eduhk.hk/english_pronunciation/index.php/2-2-formants-of-vowels/. engh, dwayne. 2013. why use music in english language learning? a survey of the literature. english language teaching 6.113-27. online: https://files.eric.ed.gov/fulltext/ej1076582.pdf irene, laurel, and harris, david. 2022. filtered listening of vocal regions. voice science works. online: https://www.voicescienceworks.org/filtered-listening-and-vocal-regions.html. nettl, bruno. 2005. the study of ethnomusicology: thirty-one issues and concepts 2.51. champaign: university of illinois press. otieno, mark owuor. 2017. what languages are spoken in hong kong? world atlas. online: https://www.worldatlas.com/articles/what-languages-are-spoken-in-hong-kong.html. piché, jean (ed.) 1994-97. table iii: formant values. the csound manual (version 3.48): a manual for the audio processing system and supporting programs with tutorials. massachusetts institute of technology. online: https://www.classes.cs.uchicago.edu/archive/1999/spring/cs295/computing_resources/ csound/csmanual3.48b1.html/appendices/table3.html. soundbridge. 2019. using formants to synthesize vowel sounds. soundbridge blog, 7 july 2019. online: https://soundbridge.io/formants-vowel-sounds/. https://altissia.org/music-as-an-effective-tool-for-learning-languages/ https://corpus.eduhk.hk/english_pronunciation/index.php/2-2-formants-of-vowels/ https://files.eric.ed.gov/fulltext/ej1076582.pdf https://www.voicescienceworks.org/filtered-listening-and-vocal-regions.html https://www.worldatlas.com/articles/what-languages-are-spoken-in-hong-kong.html https://www.classes.cs.uchicago.edu/archive/1999/spring/cs295/computing_resources/csound/csmanual3.48b1.html/appendices/table3.html https://www.classes.cs.uchicago.edu/archive/1999/spring/cs295/computing_resources/csound/csmanual3.48b1.html/appendices/table3.html https://soundbridge.io/formants-vowel-sounds/ endnotes 1 version 6.1.16 of praat software, developed by paul boersma and david weenink, was used for this study. 2 one of the sources used in the creation of this study’s digitalized sounds, eduhk (2021), hails from an academic establishment in hong kong, where according to world atlas (2017), the most spoken language is cantonese. despite this statistic, all the data reported on the establishment’s website appears to be taken from american english speakers (eduhk 2021). as this is the dialect around which the present study focuses, the author determined eduhk to still be a credible source for retrieving phonetic information. 3 all the participants except for one completed the study activity over zoom; differences in modality and any impacts they might have had on aural perception were not accounted for in this study. oral explanations in university teaching: the role of projector constructions in spoken german oral explanations in university teaching: the role of projector constructions in spoken german anna saller university of colorado boulder this study explores the nature of oral explanations in german university teaching and focuses particularly on projections, which are a widely-used feature in order to guide the students’ attention. projector constructions facilitate the production and organization of complex statements and allow the speaker to draw the students’ attention to crucial pieces of information. furthermore, projections facilitate the drawing of inferences for the listener: the projector phrase opens up semantic and syntactic slots that need to be filled, so that the number of possible contents to follow is restricted. in german, the placement of the conjugated verb in the projected unit plays a crucial role for this construction. keywords: spoken language, german, oral explanation, projector construction, construction grammar 1. relevance explaining effectively is a key component in many social spheres. however, oral explanations have hardly been subject to linguistic studies so far. because of the high impact of the speech act theory in the late 1970s and the early 1980s, explaining was examined as an isolated speech act type only (cf. spreckels 2009:1). the processual nature of explaining was neglected, and the explanation as a final product was in focus instead (cf. kiel 1999:16). the few theories about explanations and explaining that were developed often proclaim an idealized type of explaining (cf. kiel 1999:16), without considering empirical material, in order to investigate thematic or context specific deviations. a study about the form of explanatory sequences considering situation and context is still a desideratum. after giving a brief overview of the theoretical basics of explanations from a semantic point of view, i will discuss some results of a study which focused on linguistic patterns of oral explanations in university teaching, and specifically focusing on projector constructions. while one could argue that looking at explanations without considering turns is an unsatisfactory approach for interactional linguistics, explaining and understanding go hand in hand and take place in an interactive situation in which both the speaker and listeners are physically present (cf. ehlich 2009:16). furthermore, face-to-face communication comprises not only verbal and physical aspects, but is also based on reception, knowledge, and inferences (cf. fiehler 2015:373). since the term projector construction (pc) is borrowed from the theoretical framework of 1 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 construction grammar (cxg), i will briefly explain why the cxg perspective is an appropriate one when talking about spoken language, and what a pc looks like. afterwards, i will show how pcs manifest themselves in oral explanations in university teaching, how pcs in german explanations are significant, and how they fulfill an important function in guiding the students’ attention to crucial parts in an explanatory sequence. 2. the basics of explaining the verb erklären (‘explain’) is used in a very broad sense in everyday conversation. klein subdivided the concept of explaining into three semantic classes: erklären-was (‘explainingwhat’), erklären-wie (‘explaining-how’), and erklären-warum (‘explaining-why’). explaining for instance what an idiom means is explaining-what, if someone explains how to use a computer program, it is explaining-how, and explaining why glaciers are melting is explainingwhy (cf. klein 2009:25). the common goal of all these types of explaining is ensuring comprehension. the classification into those three types is based on the question one can use to ask for the explanandum (the phenomenon explained). the what type also comprises questions starting with who or which and is targeted at characteristic traits of the explanandum. the how type asks for the modality of processes, and the why type gives a cause or reason for something (cf. klein 2009:26). in german, the verb erklären is often used synonymously to erläutern (‘elucidate’) and begründen (‘give reasons’). those terms, however, can be semantically distinguished: the communicative function of erläutern (unless that of erklären) is not the systematic creation of knowledge, but adding supplementary knowledge that is necessary in the specific context of action (cf. morek 2012:31). therefore, erläutern is rather reparative in nature. begründen, too, does not primarily create new knowledge; instead, it updates and rearranges already existing elements of knowledge (cf. morek 2012:31). however, it is difficult to draw a clear line between begründen and erklären-warum (in klein’s semantic distinction), because both rearrange knowledge by giving a reason or causal relation. from a receptive perspective, the explanation process can be divided into three cognitive processes: analysis, synthesis, and syncrisis (from greek syn ‘together’ and krinein ‘choose, arrange) (cf. figure 1). during analysis, the explanandum is split into its elements and they are put into a certain relation. during synthesis, the explanandum is inserted into a superordinate complex 2 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 of knowledge. the synthesis is a central criterium, because it distinguishes explaining from other related speech acts. however, it is the syncrisis that makes an explanation successful, since here the new knowledge is anchored in the recipient’s network of knowledge (cf. hohenstein 2006:87). figure 1. the three cognitive processes of reception during an explanation (cf. hohenstein 2006:87ff.) 3. methodology in 2016, i recorded six classes at a university (regensburg, germany) in order to analyze oral explanation patterns in german university teaching. from these recordings (three in german linguistics and three in german literature classes), i isolated explanatory sequences and transcribed them based on gat1 conventions with some minor variations (contrary to most conventions, i used regular punctuation, since this facilitated the recognition and analysis of the pcs examined here). the individual sequences varied between 4:24 and 8:45 minutes in length, since the teachers varied in speaking rate and pause duration. the final corpus of the extracted explanatory sequences was 42:39 minutes in length. all classes were introductory classes (no lectures) in which the explanation of crucial contents for the particular fields was essential. the teachers recorded were balanced in ages (in their 30s, 40s, 50s and 60s), sexes, and academic status (phd, post-doctoral, and tenure-track positions). 4. spoken language and the construction grammar perspective the study aimed at investigating which strategies teachers employ in order to create cohesion in oral explanations, and whether there are differences between those strategies in linguistics and in literary studies classes. the focus was on a structure which is referred to as projection or projector construction in construction grammar and in interactional linguistics. oral 1 gat = gesprächsanalytisches transkriptionssystem (transcription system for analyses of spoken interactions); developed by selting et al. 1998. analysis splitting the explanandum into its elements synthesis explanandum is inserted into a superordinate complex of knowledge syncrisis new knowledge is anchored in the recipient’s network of knowledge 3 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 explanations take place in a face-to-face situation, which means that the sequences should be analyzed from an interactional perspective. a purely structural perspective works for written language, but not for the structure of spoken language. language routines and grammar evolve by daily language use – due to grammaticalization, grammar changes all the time in our daily interactions and is never solidified (cf. ford et al. 2003:119), which is why traditional grammars are not suitable for the description of spoken interaction. for instance, there are three assumptions in traditional grammar theories that do not work for spoken interactions: i. clause assumption: a clause is a complete syntactic unit that expresses one proposition and that consists of at least a subject and a predicate. ii. formality assumption: syntactic rules are formal, abstract, and generally valid. they apply for all elements of the respective grammatical category (e.g. word class, clause type etc.) or for respective grammatical relations (syntactic functions etc.) iii. compositionality assumption: the meaning of a phrase or clause is compositional. this means that the meaning of a phrase or clause is made up by the sum of the lexical meanings of its word and the syntactic structure in combination (cf. deppermann 2006:44). construction grammar (cxg) developed out of the recognition that those premises are inadequate for spoken language (cf. deppermann 2006:47) and is therefore increasingly used for interactional purposes. cxg is a collective term for a number of theoretical conceptions about grammar that developed under mutual influence (e.g. fillmore et al. 1988; kay 1997, 2002; kay/fillmore 1999; goldberg 1995; langacker 1987; croft 2001)2. cxg views a construction as a form-meaning pairing in which the meaning is not compositional (cf. michaelis 2006:73). the formal part of the construction is the formal realization in a scheme or syntactic pattern and comprises phonological patterns like prosody and intonation in addition to morphosyntactic patterns. the meaning part consists of both semantic and pragmatic meaning. since each of the 2 the term construction grammar (cxg) was coined by charles fillmore and paul kay (fillmore et al. 1988; kay 1997; kay/fillmore 1999). they studied particularly idiomatic phenomena of single languages. adele goldberg (1995) tied in with lakoff’s cognitive linguistics (lakoff 1987) and his cognitive view on categorization. ronald langacker’s cognitive grammar (langacker 1987) is a comprehensive language theory that describes general cognitive and symbolic principles. finally, william croft developed langacker’s approach further to a radical construction grammar (croft 2001), according to which syntactic knowledge is exclusively represented by grammatical constructions. 4 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 aforementioned elements are crucial to understanding the use of projections in oral explanations, a perspective that is aligned with basic assumptions of cxg is reasonable here. however, while cxg has a strong cognitive perspective, conversation analysis (ca) focuses on interaction, which is why most studies in interactional linguistics have focused on semantic and pragmatic aspects so far (cf. deppermann 2006:59), but not in connection to structural features. furthermore, it has not yet been investigated how certain constructions are actually used in particular contexts or genres of spoken interaction (cf. günthner 2006:98). thus, this study explores pcs in the context of university teaching. 5. projections in oral explanations a projection is a structure which consists of two parts: the first part (a) is the projector phrase which projects the second part (b) semantically and phonetically as well as syntactically (cf. günthner 2008:86). in order to understand why those projections are significant here, one must consider that in german, the position of the conjugated verb depends on whether it occurs in a main or in a subordinate clause. in a main clause, the conjugated verb (which might be an auxiliary in perfect tense or in passive voice, or a modal verb) is in second position, regardless of what is in first position (unlike english, where the verb usually follows the subject or agent). in a subordinate clause, the conjugated verb is in the last position. the projector phrase a, which announces b, is placed in square brackets. the main verb in a is underlined, because it is crucial for the projection – it requires an obligatory complement which is realized in b as a complement clause. even if b is supposed to be a “subordinate” clause according to traditional grammar, it has a v2 pattern. (1) [a1: also, dass sie jetzt nicht nur sagen:] [‘so, that you don’t just say now’:] [a2: gut, ich hab gelernt:] b: es gibt x (.) laute und diese x laute, die ändern sich irgendwie, und das hab ich auswendig gelernt. (ling) [‘good, i have learnt’:] ‘there are x (.) sounds und those x sounds, they change somehow, and i have learnt that by heart.’ (2) [a: ich hab noch gelernt:] b: da gibt es eine sogenannte auslautverhärtung=ja. (lit) [‘i have still learnt’:] ‘there is a so-called final-obstruent devoicing=ya.’ about a century ago, the b parts as shown above used to be called “uneingeleiteter nebensatz” (‘unintroduced subordinate clause’) (behaghel 1928), because the formulation in 5 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 written language for b would have been introduced by a subordinating conjunction (here: dass) with the verb in final position (compare 3b, 4b): (3) a. [a2: gut, ich hab gelernt:] b: es gibt x (.) laute … b. gut, ich habe gelernt, dass es x laute gibt … (4) a. [a: ich hab noch gelernt:] b: da gibt es eine sogenannte auslautverhärtung=ja. b. ich habe noch gelernt, dass es da eine sogenannte auslautverhärtung gibt. in our examples, there is neither a subordinating conjunction nor is the verb in final position. and it is actually the b part of the construction that carries the relevant information (the central proposition), so one could argue that semantically part b is not subordinate at all. this reveals the flaws in the notion of “main” and “subordinate” clauses, which are very much aligned to traditional grammar views based on written language. since b has the same syntactic structure as a main clause in german, it was later called “abhängiger hauptsatz” (‘dependent main clause’) (auer 1998), which seems to be contradictory in itself. i argue that b is not dependent on part a, but rather the projector phrase raises the awareness of the recipient for the following b part as it carries crucial information. the projector phrase is syntactically, semantically, and intonationally incomplete and opens up slots. the incompleteness in the projector phrase draws the attention of the recipients to exactly those slots that are filled by the b part. thus, the projector phrase plays an important role on a metacommunicative level, but it is does not contribute to the proposition conveyed in b. the basic units of spoken language are rarely complete sentences as presented in traditional grammar based on written language. instead, the basic units are intonationally and semantically coherent and functional units that usually just carry one new piece of information (cf. tomasello 2003:4). auer found that in constructions that consist of two parts (like projector constructions (pcs)), the transition between a and b is usually marked by a pause or an increase in pitch (cf. auer 1997:61). this increase in pitch along with rising intonation indicates phonetically that the construction is not complete yet, and thus, it secures the right to continue speaking. 6 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 despite the fact that pcs clearly consist of two parts (which is a typical feature of spoken language; cf. fiehler 2015:3913), the use of terms like “dependent” and “independent” or “subordinate” seems to be inappropriate for spoken language. first of all, the notion of dependence is one of a hierarchical organization of syntax which is based on traditional views of grammar for written language, and second, utterances are constantly under construction. the notion of a hierarchical structure makes sense for written language production where a whole sentence is planned beforehand. in spoken interaction, however, the planning of utterances happens in every moment of speaking. this bottom-up approach to speech production is supported and confirmed by psycholinguistic findings (cf. rickford et al. 2010:55). the temporally undelayed production and reception of utterances during spoken interaction is also called online syntax or incremental syntax (cf. auer 2000, 2007). the immediate production during speaking requires different strategies than for writing. thus, the two basic operations for spoken language are projections and retractions (cf. auer 2000:47). the use of the type of pc discussed here4 allows the speaker to (i) conceptualize the main information, (ii) secure his right to continue speaking, and (iii) draw the attention to the slots that are opened up by the projector phrase. the induced control of the recipients’ expectations works on a syntactic, semantic, and intonational level. there were hardly any differences in the number of pcs between explanations in linguistics and in literary studies in my corpus: there were 23 pcs in linguistic explanations, and 27 in literary studies explanations (which makes 50 pcs in total). those 50 pcs allocated to 6 recordings means 8.33 pcs on average per explanatory sequence. the average length of an explanatory sequence 3 projector constructions are very similar to what fiehler calls operator-scopus structre: this is a spoken unit consisting of two parts. the scopus is a full proposition, and the operator is a preceding unit that refers to the scopus and acts as an “instruction manual” for the recipient on a metacommunicative level, which means that it gives the recipient a hint on how to understand the following utterance. (cf. fiehler 2015:386, 391). 4 there are many different types of projector constructions or projections. they can be lexically more solidified like die sache ist constructions in german (cf. günthner 2008), but they can also be purely structural like ‘pseudocleft’ constructions (cf. ibid). in german, an inflected adjective only occurs within a np and thus projects a noun; a preposition requires and projects a certain case; a verb in a certain context projects the number of complements depending on its valency). german is a language which is rich in projections compared to japanese which is poor in projections but uses other strategies in conversation (cf. auer 2007:4; ford et al. 2003:130f.). there is a general tendency that languages that employ more synthetic strategies (inflection) are richer in projections, and languages that employ more analytic strategies are poorer in projections. 7 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 was 6:35 minutes, which means that teachers used on average 1.31 pcs per minute of explanatory sequence (or at least 1 pc per minute for explaining). 5.1. form of projected units as discussed above, the crucial element of this construction is that part b does not show the typical features of a “subordinate” clause – neither structurally, nor semantically. looking at the projected units more closely, one finds that they do not necessarily show a v2 pattern as well, but express questions (examples 5 and 6) and units with no explicit verb at all (examples 7 and 8): (5) [a: und jetzt müssen sie wissen:] b: wie nennt man diese zwei gruppen der mittelhochdeutschen diphthonge? (ling) [‘ad now you have to know’:] ‘how are these two groups of middle high german diphthongs called?’ (6) [a: sie schauen sich sozusagen das ende an:] b: wie kuckt das ende eines verses aus? (lit) [‘yu look, as it were, to the end’:] how does the end of the verse look like?‘ (7) [a: bei [ts] haben wir die geschichte auch:] b: beide male dental, (-) erst plosiv, dann frikativ. (ling) [‘e have the story with [ts], too’:] ‘both times dental, first plosive, then fricative.’ (8) [a: ich habe das deswegen so ausführlich gemacht, um ihnen einen begriff nochmal nahezubringen, den sie eh schon mal letzte sitzung kurz in den raum gestellt haben:] b: die kadenz. (lit) [‘i made this so detailed with the purpose to confront you with a term again which you already mentioned (literally: ‘placed into the room’) once last class’:] ‘cadence.’ this makes the term “abhängiger hauptsatz” seem even more unsuitable, particularly for the questions. however, in constructions like 7 and 8, the b parts are indeed dependent on the projector phrase, and the whole construction works differently: here, it is not the main verb in the projector phrase that opens up a syntactic slot which needs to be filled, but it is a noun phrase (np). in order to guide the recipients’ attention, the speaker uses a substitution for b in the form of a np in the projector phrase. in example 7, the geschichte (‘story’) refers to the earlier posed question of what an affricative is. this question was answered for [pf] before (as a transitional sound from a plosive to a fricative), and now the story is the same for [ts]: ‘both times dental, first plosive, then fricative’. in example 8, kadenz (‘cadence’) is substituted by einen begriff (‘a term’) in the projector phrase. in those examples without an explicit verb in b, one can also speak of 8 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 retraction, because b is reduced to the greatest degree, but at the same time refers back to a crucial semantic slot that was opened up in a. those examples show how projection and retraction can go hand in hand with each other. figure 2 shows the distribution of v2 pattern, question pattern, and verbless pattern in the projected units within the corpus. figure 2. verb position patterns in the projected units (b) in total, 24 out of 50 units showed v2 pattern (48%), 11 out of 50 were questions (22%), and 15 out of 50 did not have an overt predicate (30%). all teachers use pcs with v2 pattern. 5 out of 6 teachers use verbless structures, and 3 out of 6 teachers use questions as well. questions in projected units are therefore the least frequently used forms in projected units, and all questions are question-word interrogatives. there is only one speaker (ling explanation) who uses verbless units in b more often than the v2 pattern, which might be an idiosyncratic feature. those retractive constructions with verbless b parts have a strong binding force, because a and b are strongly dependent on each other: a does not make sense without b, because there is a crucial semantic gap in it – and the semantic gap, which is b, does not make sense on its own as well, because it is the “missing piece” in the projector phrase. a few more examples shall illustrate that: (9) [a: manchmal werden die auch tenues genannt, deswegen dieser begriff hier:] b: der tenuesverschiebung. (ling) [‘sometimes, they are also called tenues, therefore this term here’: ] ‘tenues shift.’ (10) [a: dann ein nächstes beispiel:] b: em (-) barocke emblematik. (lit) [‘then another example’: ] ‘baroque emblematics.’ 3 7 0 2 2 1 4 2 1 3 6 8 5 0 0 0 1 5 0 2 4 6 8 10 ling 1 ling 2 ling 3 lit 4 lit 5 lit 6 question v2 no overt verb 9 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 5.2. form of projector phrases as is typical for pcs, which consist of a projector phrase a and a projected unit b, the conjugated verb in part b is never in final position. however, the projector phrase (which, according to traditional grammar views is a “main” clause) can have the verb in final position. figure 3. verb position patterns in the projector phrases (a) figure 3 shows that the v2 pattern is still the most frequent one in the projector phrase (32 out 50, which makes up 64%), followed by verbless units (10 out of 50, which makes up 20%), and units with the verb in final position occur in 16% of all cases (8 out of 50). in those cases with the verb in final position, the projector phrase was either introduced by a subordinating conjunction (dass ‘that’, wenn ‘when’/ ‘if’, weil ‘because’, was ‘what’[relativizer]) or it was an infinitive construction (um … zu ‘(in order) to’). examples 11 and 12 illustrate this with wenn: (11) [a: wenn sie jetzt sagen:] b: wie(.)so schreiben die da ein und sprechen ein [ı]? [– ja, keine ahnung, das machen wir auch.] (ling) [‘if you say now’: ] ‘why do they write an and speak an [ı]?’ [ – ‘ya, no idea; we do that, too.’] (12) [a: wenn sie irgendwo sehen:] b: aha, da ist ein [p], das betroffen ist, [dann heisst das erstmal noch !gar! nichts.] (ling) [‘when you see somewhere’: ] ‘ah, there is a [p] which is affected,’ [‘that does not mean anything provisionally.] looking at the construction as a whole, it turns out that not all constructions have the form [[v2:]v2], but it is the most common form. the verb in final position occurs in 64% of all cases in 3 3 0 1 2 1 7 3 1 4 6 11 2 4 0 0 1 1 0 2 4 6 8 10 12 ling 1 ling 2 ling 3 lit 4 lit 5 lit 6 v in final position v2 no overt verb 10 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 part a and in 48% of all cases in b. in example 12, b is followed by another unit in square brackets, because whenever there is a unit introduced with wenn, it is followed by another unit introduced by dann. the wenn…-dann… (‘when…-then…’) construction is a typical structure that is found very often in explanations (cf. morek 2012:29), since it expresses a regularity in a certain condition. 5.3. lexical, syntactic, and pragmatic conditions for projections the observation that v2 patterns following a “main clause” occur much more frequently in spoken than in written german is in fact not new: auer (1998) compared the freiburg corpus, created specifically to study the syntax of spoken language, to the bzk (bonn-kölnerzeitungskorpus ‘bonn-cologne newspaper corpus’ for written language), and found that the v2 pattern ‘dependent main clause’ occurs significantly more frequently in spoken german, and that it correlates with specific semantic fields: they are often used after verba sentiendi and verba dicendi (verbs of reception and communication), for instance hören (‘hear’), denken (‘think’), or sagen (‘say’). however, v2 patterns in b are usually not compatible with a parts that use verbs of wanting and causing, for instance veranlassen (‘to cause’), wollen (‘to want’), or verhindern (‘to prevent’) (cf. auer 1998:8). indeed, the pcs in the analyzed data mostly used verbs in the projector phrase that express a form of thinking, remembering, communicating etc. a ‘dependent main clause’ is also not possible after negation. example 13 (constructed) illustrates that syntactic negation is not compatible with a following v2 pattern. 13a is a positive statement (‘i think’) that works perfectly fine with a v2 pattern, but 13b, which is just the negation of 13a, does not work with this pattern. (13) a. [ich denke: ] es wird morgen regnen. [‘i think’: ] ‘it will rain tomorrow.’] b. *? [ich denke nicht: ] es wird morgen regnen. *? [‘i don’t think’: ] ‘it will rain tomorrow.’] however, it is unclear whether it is the syntactic construction that makes 13b sound awkward, or whether it is the negative semantics. example 14 (constructed) indicates that it is more likely that the negative semantics is not compatible with the v2 pattern, because here there is no syntactic 11 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 negation, but rather verbs which have an inherently negative semantics (‘to doubt’, ‘to be false’). in those cases, the subordinating conjunction dass with verb in final position is required. (14) a. [ich hoffe: ] es gibt ein leben nach dem tod. [‘i hope’: ] ‘there is a life after death.’] b. *? [ich bezweifle: ] es gibt ein leben nach dem tod. *? [‘i doubt’: ] ‘there is a life after death.’] (15) a. [es ist wahr: ] in seattle regnet es viel. [‘it is true’: ] ‘it rains a lot in seattle.’] b. *? [es ist falsch: ] in seattle regnet es viel. *? [‘it is false’: ] ‘it rains a lot in seattle.’] interestingly, comparing german to english, it seems to be more common in english as well to insert the subordinating conjunction that after projections with negative semantics (i.e. i doubt that there is a life after death; it is false that it rains a lot in seattle). furthermore, there are important pragmatic reasons which invoke projected units with v2 patterns. units that are introduced by subordinating conjunctions are more likely to contain already familiar content, and are thus pushed into the background, which means that the first part of the construction is in the focus (cf. auer 1998:11). conversely, if the second part is not introduced by a subordinating conjunction but shows a v2 pattern, the attention is drawn to this second part (cf. ibid). this might even explain why utterances with a negative semantics often occur with subordinating conjunctions: we do not negate random facts that we consider to be false, but only those ones of which we can assume that out communication partners are already familiar with, and that are relevant to them (cf. auer 1998:11). in other words: the new and relevant information is the negation in the a part itself, and the thing being negated in b is already familiar information, that is moved away from the focus. similarly, if the first part of the construction already implies a certain knowledge, attitude or mental action of the speaker (e.g. staunen ‘to wonder’, gut finden ‘to like’, verzeihen ‘to forgive, sich wundern ‘to be surprised’), it is more likely to be followed by a unit with a subordinating conjunction and verb in final position. this is because the relevant information is implied in the first part and the first part is in focus, not the second one, as example 16 shows. 12 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 (16) a. ich wundere mich, dass du hier bist. ‘i am surprised that you are here.’ b. *? [ich wundere mich: ] du bist hier. *? [‘i am surprised’: ] ‘you are here.’ the assumption that the v2 pattern is used to guide the recipient’s attention is also backed by the fact that a construction of the form [[v2:]v2] does not allow the second part to be moved in front of the first part5, which is possible if the same utterance has the form [[v2:]vfinal]: (17) a. ich erkenne, dass die zeit reif ist. ‘i realize that the time is ripe.’ a’. dass die zeit reif ist, erkenne ich.] b. [ich erkenne: ] die zeit ist reif. [‘i realize’: ] ‘the time is ripe.’ b’. *?die zeit ist reif, erkenne ich. 17a’ illustrates that the clause which is introduced by a subordinating conjunction and has the verb in final position can be moved into the prefield (the slot in front of the conjugated verb), whereas for 17b’ with the clause in v2 structure, this is impossible. the fact that the structure with subordinating conjunction and vfinal can be topicalized means that there is less focus on this part, since the topic/theme carries usually familiar or less relevant information. the v2 pattern, however, can not be topicalized, which is an indicator that the information is considered to be important, and that this structure is used to shift the focus on this part of the construction. this observation fosters the hypothesis that pcs of the type with ‘dependent main clause’ are used to draw the recipients’ attention to the second part of the construction (as well as the fact that the first part builds up semantic and syntactic expectations in the hearer by the slots it opens up). 5 there is the tradition in german linguistics to describe constituent order in a sentence with the theory of topological fields. since german is considered to be a v2 language, the syntactic phrase in front of the conjugated verb is called vorfeld (‘prefield’), the phrase(s) between the first and the second (if applicable) part of the predicate (which can consist of linker and rechter satzklammer (‘left and right sentence brace’) in perfect tense, constructions with modal verbs, passive etc.) is called mittelfeld (‘middle field’), and everything following the right sentence brace is called nachfeld (‘afterfield’) (usually subordinate clauses in written language, and obliques in spoken language if there were too many obliques and complements in the middle field). 13 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 5.4. indirect change of perspective the interactional nature of oral explanations also becomes obvious in what i call “indirect change of perspective”. this means that the teacher formulates an utterance from the perspective of the recipients, which is discernable in the use of the pronouns. by doing that, the teacher leaves the role of disseminator of knowledge and assumes the role of the students. those changes of perspective very often occur in b parts of pcs. example 18 illustrates that well: (18) [a1: also, dass sie jetzt nicht nur sagen:] [‘so, that you don’t just say now’:] [a2: gut, ich hab gelernt:] b: es gibt x (.) laute und diese x laute, die ändern sich irgendwie, und das hab ich auswendig gelernt. (ling) [‘good, i have learnt’:] ‘there are x (.) sounds und those x sounds, they change somehow, and i have learnt that by heart.’ in a1, the students are directly addressed as sie’/you’, which explicitly excludes the speaker. a1 is here the projector phrase for [[a2:]b], which is an indirect change of perspective because it articulates what the students are supposed to think. in [[a2:]b], the pronouns switches from sie/’you’ to ich/’i’. within the indirect change of perspective, a2 serves as a projector phrase for b as well, so that the whole structure is [[a1:]a2:]b]. the personal pronoun ich (‘i’) does not refer to the teachers themselves, but to each individual student from their perspective. additionally, all the questions that occur in b parts (which are 22%, see 5.1) are changes of perspective, because they express what students are supposed to ask themselves. the previously provided examples 5 and 11, reproduced below, illustrate this well: (5) [a: und jetzt müssen sie wissen:] b: wie nennt man diese zwei gruppen der mittelhochdeutschen diphthonge? (ling) [‘and now you have to know’:] ‘how are these two groups of middle high german diphthongs called?’ (11) [a: wenn sie jetzt sagen:] b: wie(.)so schreiben die da ein und sprechen ein [ı]? [– ja, keine ahnung, das machen wir auch.] (ling) [‘if you say now’: ] ‘why do they write and speak [ı]?’ [ – ‘ya, no idea; we do that, too.’] articulating potential or supposed thoughts of the students can help them process the explained content and create mental coherence. the whole corpus showed a total of 18 indirect changes of 14 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 perspective, of which 15 were used in projected unit, meaning that in 83.3% of all changes of perspective it follows a projector phrase. by doing this, the mental processing is facilitated in two ways: first, by opening up slots and thus guiding the recipients’ attention, and second, by assuming the role of the student and articulating content from their perspective. figure 4. indirect change of perspective in projected unit within pcs 6. conclusion in spoken language, coherence is created differently in both production and perception than in written language, and the speakers use different cohesive strategies. in oral explanations in university teaching, projector constructions are used frequently, and the view of this structure as a construction includes not only the syntactic structure, but also semantics, pragmatics, and prosody which interact closely with one another. even if pcs are analyzed in their two parts here, the two units form one construction, and the pc can only display its effect (to guide one’s attention) in spoken interaction as a whole. this is why the view of german syntax as consisting of matrix and subordinate clauses is not unproblematic when talking about spoken language. the reasons why pcs are used so frequently in oral explanations are to be found in their cognitive and interactive functions: for the speaker, it is easier to produce and to organize complex statements, especially given the fact that speech is produced online, and conceptualization is an incremental and ongoing process (cf. auer 2000). by using projections, the speakers secure their right to continue speaking. from the receptive perspective, pcs facilitate the drawing of inferences: the projector phrase opens up specific slots that need to be filled, so that the number of possible contents to follow is restricted (cf. günthner 2008:108). in university teaching, pcs occur in various forms. the [[v2:]v2] construction is the most common one, but there are also vfinal as well as verbless projector phrases (a). the projected units 6 1 0 0 1 7 0 2 0 0 0 1 ling 1 ling 2 ling 3 lit 4 lit 5 lit 6 yes no 15 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 (b) can be in question form as well, but they do not have a vfinal pattern (if they do, they are not considered a pc, but a hierarchical structure following the rules of written language). when b occurs as a verbless unit, it usually does not fill a syntactic gap opened up by the valency of the main verb in a, but it substitutes a less specific np in the projector phrase that is used to draw the recipients’ attention to b. another interesting observation is that the teachers sometimes form utterances from the students’ perspective, which raises the potential for identification in them. in more than 80% of those cases, the change of perspective occurs in part b within a pc. the suggestion that pcs are used to guide the students’ attention is supported by some restrictions to pcs: part a cannot be negated, because a negation is always in focus, which is contradictory to a pc whose aim is to shift the focus to part b. part b, in turn, cannot be topicalized, because topicalization would lead to reduction in focus. if b were phrased as a “subordinate clause” with a vfinal pattern, however, it could be topicalized, but then, it is out of focus. this shows that the verb placement in german is crucial in order to either put units in focus or to push them into the background. this study shows how pcs are systematically used in university teaching to guide the students’ attention to crucial pieces of information in part b, and how oral explanations in german exploit a different strategy with regard to the syntactic structure than written german does. references auer, peter. 2007. syntax als prozess. gespräch als prozess, ed. by heiko hausendorf, 95–124. tübingen: narr. online: http://www.inlist.uni-bayreuth.de/issues/41/inlist41.pdf, 1–35. auer, peter. 2000. on-line syntax – oder: was es bedeuten könnte, die zeitlichkeit der mündlichen sprache ernst zu nehmen. sprache und literatur 85. 43–56. auer, peter. 1998. zwischen parataxe und hypotaxe: abhängige hauptsätze im gesprochenen und geschriebenen deutsch. zeitschrift für germanistische linguistik 26. 284–307. auer, peter. 1997. formen und funktionen der vor-vorfeldbesetzung im gesprochenen deutsch. syntax des gesprochenen deutsch, ed. by peter schlobinski, 55–92. opladen: westdeutscher verlag. behaghel, otto. 1928. deutsche syntax. eine geschichtliche darstellung. band iii: die satzgebilde. heidelberg: winter. croft, william. 2001. radical construction grammar. oxford: oxford university press. deppermann, arnulf. 2006. construction grammar – eine grammatik für die interaktion? grammatik und interaktion. untersuchungen zum zusammenhang von grammatischen 16 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 strukturen und gesprächsprozessen, ed. by arnulf deppermann, reinhard fiehler and thomas spranz-fogasy, 43–66. radolfzell: verlag für gesprächsforschung. ehlich, konrad. 2009. erkären verstehen – erklären und verstehen. erklären. gesprächsanalytische und fachdidaktische perspektiven, ed. by rüdiger vogt, 11-24. tübingen: stauffenburg. fiehler, reinhard. 2015. syntaktische phänomene in der gesprochenen sprache. handbuch satz, äußerung, schema (= handbuch sprachwissen 4), ed. by christa dürscheid and hans georg, 370–395. berlin et al.: de gruyter. fillmore, charles; paul kay; mary catherine o'connor. 1988. regularity and idiomaticity in grammatical constructions: the case of let alone. language 64, vol. 3. 501–538. ford, cecilia; barbara fox; sandra a. thompson. 2003. social interaction and grammar. the new psychology of language: cognitive and functional approaches to language structure, vol. 2, ed. by michael tomasello, 119–143. mahwah, nj: routledge. goldberg, adele. 1995. constructions. chicago, il: university of chicago press. günthner, susanne. 2008. projektorkonstruktionen im gespräch: pseudoclefts, die sache istkonstruktionen und extrapositionen mit es. gesprächsforschung 9. 86–114. günthner, susanne. 2006. grammatische analysen der kommunikativen praxis – ‚dichte konstruktionen‘ in der interaktion. grammatik und interaktion. untersuchungen zum zusammenhang von grammatischen strukturen und gesprächsprozessen, ed. by arnulf deppermann, reinhard fiehler and thomas spranz-fogasy, 95–122. radolfzell: verlag für gesprächsforschung. hohenstein, christiane. 2006. erklärendes handeln im wissenschaftlichen vortrag. ein vergleich des deutschen mit dem japanischen. münchen: iudicium. kay, paul, and charles fillmore. 1999. grammatical constructions and linguistic generalizations: the what's x doing y? construction. language 75. 1–33. kay, paul. 1997. words and the grammar of context. stanford, ca: csli. kiel, ewald. 1999. erklären als didaktisches handeln. würzburg: ergon. klein, josef. 2009. erklären-was, erklären-wie, erklären-warum. typologie und komplexität zentraler akte der welterschließung. erklären. gesprächsanalytische und fachdidaktische perspektiven, ed. by rüdiger vogt, 25–36. tübingen: stauffenburg. 17 saller: projector constructions in oral explanations in german university teaching published by cu scholar, 2019 lakoff, george a. 1987. women, fire, and dangerous things. what categories reveal about the mind. chicago, il: university of chicago press. langacker, ronald w. 1987. foundations of cognitive grammar, vol. 1. stanford, ca: stanford university press. michaelis, laura a. 2006. construction grammar. the encyclopedia of language and linguistics, 2nd edition, vol. 3, ed. by k. brown. 73–84. oxford: elsevier. morek, miriam. 2012. kinder erklären. interaktionen in familie und unterricht im vergleich. tübingen: stauffenburg. rickheit, gert; sabine weiss; hans-jürgen eikmeyer. 2010. kognitive linguistik. theorien, modelle, methoden. tübingen: francke. selting, margret; peter auer; birgit barden et al. 1998. gesprächsanalytisches transkriptionssystem (gat). linguistische berichte 173. 91–122. spreckels, janet. 2009. erklären im kontext – neue perspektiven. erklären im kontext. neue perspektiven aus der gesprächsund unterrichtsforschung. ed. by janet spreckels, 1–10. baltmannsweiler: schneider hohengehren. tomasello, michael (ed.) 2003. the new psychology of language: cognitive and functional approaches to language structure, vol. 2. mahwah, nj: routledge. 18 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/4 doi: http://dx.doi.org/10.33011/cril.24.1.4 colorado research in linguistics 6-2019 oral explanations in university teaching: the role of projector constructions in spoken german anna saller recommended citation microsoft word saller-cril2019-final.docx microsoft word six-cril2021-proof_final.docx 1 report on vocal fry in interactional contexts creaky voice and pitch as affected by age and gender of speaker sophia six university of colorado boulder previous research on vocal fry has been largely one-sided, investigating its role as a vocal phenomenon for young women specifically, and focusing on countering or confirming harmful stereotypes surrounding vocal fry as perpetuated by the media. this study attempts to expand an understanding of vocal fry by investigating its production by multiple genders and ages, as well as its role in an interactional context. in this study, participants of four age and gender groups (older women, older men, younger women, and younger men) were asked to conduct a conversation with a member of each identity in order to determine the effect of identity in an interactional role on the occurrence of vocal fry. while prior studies have implied that the use of vocal fry is most heavily dependent on the identity of the speaker, this study concludes that the identity of the participant may actually play a more important role in its occurrence. it also contextualizes the prominence of female vocal fry in relation to the frequency of vocal fry by other speakers. keywords: vocal fry, creaky voice, interaction, sociolinguistics, gender, age, sexism, phonetics 1. introduction vocal fry has long been vilified by experts from every field and casual listeners alike as sounding “less competent, less educated, less trustworthy, less attractive, and less hirable”, or even “vulgar”, “repulsive”, “mindless”, and “really annoying” (anderson & klofstad 2014; garfield 2013). many critics also uphold the persistent claim that it is damaging to one’s vocal cords. speech pathologist susan sankin infamously claimed in 2015 that “if [vocal fry] is a repetitive habit that you use over a long term … the vocal cords will show some sort of fatigue” (gross). one of the early studies on vocal fry published in the journal of voice in 2011 by wolk and abdelli-beruh also claimed that “the habitual use of fry is atypical and possibly a form of vocal abuse”. it is not difficult to notice that young women are almost exclusively the target of these criticisms, vocal fry being associated with women like kim kardashian and britney spears, and therefore, by extension, with the negative stereotypes surrounding the type of women who use it (anderson & klofstad 2014; chao & bursten 2021). in reality, while the phenomenon does occur most frequently in young women, nearly every demographic uses vocal fry, regardless of age and gender. a 2015 study by callier and podesva of stanford university found that strong creak colorado research in linguistics, volume 25 (2021) 2 patterns also occur with young men and older women, and older men use some creak but with a different pattern of intonation. a 2016 study from the journal of the acoustical society of america by sarah irons and jessica alexander found that young men may even use vocal fry more often than young women. regardless, young people, especially women, are overwhelmingly the ones criticized for it. after the journal of voice published its pivotal paper “habitual use of vocal fry in young adult female speakers” in 2011, vocal fry suddenly became a viral household term. a flurry of controversial and angry thinkpieces attacking vocal fry as an unprofessional, intolerable, and damaging “fashion choice” unique to young women overpowered the media, far too many in number to begin to list here, and the issue remains controversial to this day (reynolds 2015). in an ironic catch-22, while women with a naturally higher pitch of voice are seen as less competent and less authoritative, women who lower their voice using vocal fry are also seen as being less confident and less educated (anderson & klofstad 2014). contrary to popular belief, vocal fry is not damaging to the vocal cords, as can be proven by the fact that creaky voice is used as a contrastive linguistic feature in several languages, such as danish, burmese, kwakw’ala, and jalapa mazatec, in which a word can have a different semantic meaning depending on its vocal quality (gordon & ladefoged 2001; keating et al. 2015). for example, jalapa mazatec contrasts between modal (normal) voice, creaky voice, and breathy voice, já (modal) meaning “tree”, já̤ (breathy) meaning “he wears”, and já̰ (creaky) meaning “he carries” (gordon & ladefoged 2001). if there are entire cultures that use creak every day to contrast semantic meaning, it cannot be unhealthy or damaging, and far less a speech disorder, as some would call it. research on the occurrence of vocal fry, like this study, can serve to dispel harmful stereotypes about the phenomenon, reduce relevant sexist biases and increase its normalcy, and further understand its role in linguistics. vocal fry is a type of creaky voice, or sometimes used as another term for vocal creak. it is produced when the glottal closure is loose and relaxed at a very low pitch, allowing air to vibrate through slowly and irregularly, with audible pulses that sound like a rattle or croak. in this study, i will be measuring vocal fry by measuring the harmonics and pitch of vowels. vocal fry should correlate with a greater difference in amplitude between the first and second harmonics, and a lower fundamental frequency (keating et al. 2015). this experiment will study the effect of age and gender on vocal fry in an interactional context. i have designed interviews pairing speakers of different age and gender demographics (older women, younger women, older men, and younger report on vocal fry in interactional contexts 3 men) to determine whether a speaker’s use of vocal fry is affected by the age and gender of their conversational partner, as well as whether their own age and gender affects their vocal quality. my first hypothesis is that fry will be more common among young speakers. this is supported by existing research that finds it to be a relatively new phenomenon among the younger generation (anderson & klofstad 2014; wolk & abdelli-beruh 2011). because young women are the primary users of vocal fry, and older generations are particularly offended by its usage as an undesirable feminine trait, i suspect that there will be a difference in which gender uses creak more between generations (anderson & klofstad 2014). my second hypothesis is that among young people, women will use vocal fry more, and among older people, men will use vocal fry more, as older women and younger women will draw a larger distinction between themselves in this matter, with differing definitions of a desirable vocal quality. i believe, however, that age will be the more important factor, and thus young females will use the greatest amount of vocal fry, but will be followed by young men before older men. due to recent studies indicating a strong use of vocal fry among young men, i do not believe there will be an exceptionally large difference between young men and young women in their use of vocal fry, and there will be a bigger difference between older men and older women; older women using hardly any (irons & alexander 2016). my third hypothesis is that age of the conversational partner will be the biggest factor in determining an individual speaker’s use of vocal fry. because vocal fry has become increasingly prevalent as a social phenomenon in recent years, i hypothesize that all speakers will use more vocal fry when speaking with young partners. my fourth hypothesis is that i believe that differences due to a partner’s gender will experience a shift across generations. older speakers will use more vocal fry when speaking to women, and younger speakers will use more vocal fry when speaking to men. thus, young women speaking to young men will use the most vocal fry, and older men speaking to other older men will use the least. this inference comes from a suspicion (and vague implication made by some of the research) that older generations view vocal fry in women as unattractive due to a lower pitch being seen as more masculine, and younger generations seeing a lower (or more “masculine”) pitch as a positive trait holding more gravitas (anderson & klofstad 2015; chao & bursten 2021). thus, an older man would perhaps be more likely to use vocal fry to confirm his masculinity in distinguishing himself from the more “effeminately” higher-pitched opposite sex (and therefore an older woman would do the opposite with a male partner, but may relax around another woman), while a younger female may use it to increase her colorado research in linguistics, volume 25 (2021) 4 status among male peers (with young males also conducting themselves oppositely). i also expect there to be some differences in the amount of vocal fry used based on the position of the speech sound within an utterance, or over time in an interaction. my fifth hypothesis is that i believe that vocal fry will occur at higher levels and a greater rate at the end of a phrase or utterance, and that it will increase slightly over the course of an interaction. i will be measuring vowels, but i do not expect there to be a difference in creak dependent on vowel quality, except perhaps for more creak in mid/central vowels (schwa), as i expect vocal fry to very frequently occur on the filler word “uh”, as speakers relax their vocal cords. 2. methods in order to collect data for this project, i arranged interviews between subjects of varying ages and genders. i solicited eight participants, half of them male and half female. half were under 25 and half were over 40, in order to establish a noticeable age gap. i have changed participants’ names here in order to protect their privacy. each participant met for an interview with four other participants, one from each age and gender demographic. see figure 1. figure 1. participants & conversation pairing report on vocal fry in interactional contexts 5 most of the participants did not previously know one another. due to logistics issues, i did have to make two pairings between people that had established relationships. heidi is the mother-inlaw of tania’s sister, so the two are friendly, but have not spent enough time together to have a very close relationship. catherine and rob are a married couple, however. i was hoping to avoid pairings between people with a close personal relationship, but unfortunately i could not pair them with alternate participants for scheduling reasons. none of the other participants had previously met. all speakers are native speakers of american english and live in colorado. all identify as caucasian, but landon also has some latino background and dana has some cherokee background. tania is from washington, dana is from tennessee, seth is from maine, rob and catherine are from new york, landon is from texas, and andrew and heidi are from california. each of their idiolects reflects a mix of features reflective of their home state as well as a standard coloradan accent. participants met over zoom. i supervised and recorded sessions but remained muted during the interviews. i asked participants to turn on their cameras during the interviews. i gave participants a script ahead of time to interview one another with. see figure 2. the reason for a scripted set of interview questions was in order to a) maintain topic consistency so that i would have words or phrases to examine that would serve as minimal pairs repeated by a speaker in every session, and b) to combat the awkward silences inevitable when meeting a stranger for the first time, and personal differences in talkativeness and conversational productivity. i chose questions that would be mostly surface level and easy to talk about, but would allow speakers to briefly monologue during their answers. i instructed participants to take turns completing each overarching topic before switching roles. for example, participant 1 would ask participant 2 questions 1, 1a, 1b, and 1c, then participant 2 would ask participant 2 the same questions before they moved to the second set of questions. i encouraged participants to speak naturally and comfortably and to interject, react, or go off topic whenever it felt appropriate, and i warned them not to script or write down their answers. meetings usually lasted 15-20 minutes, although a couple of the more talkative participants went off topic and lasted 30-40 minutes. colorado research in linguistics, volume 25 (2021) 6 figure 2. interview questions with eight participants each meeting with four of the others, this culminated in 16 meetings. i recorded audio separately for each partner, so the result was 32 sound files. in order to take a small sample of the data, i segmented 13 sections of each recording, roughly equivalent to how many relevant turns they took to answer the given questions. then i began to measure utterance-initial and -final vowels. in order to avoid the complication of segmenting utterances in a given turn, i simply marked an entire multiple-utterance answer and measured the beginning and end of each answer. in this way, i was able to ensure that i was not making any errors mismarking the beginning or end of an utterance. 1. what do you do for work? (alternatively: a previous job/dream job) a. what does your job entail? b. what is the hardest part of your job? why? c. what is the most fulfilling part of your job? why? 2. did you go to college? (if not, would you have liked to? answer the rest hypothetically, with what you would like to study if you went to college) a. where did you go? b. what did you study? c. what did you find the most interesting about your studies/field of interest? 3. if salary/education/skills were not an issue, what would your dream career be? why? a. where in the world would you live if you had your dream job? b. what would you do in your free time? 4. talk about your pets! (alternatively: a favorite past pet or dream pet) a. name, species, breeds, gender, age, etc. b. what is a silly habit your pet has? c. (feel free to share photos or model them on camera if you want to!) 5. wrapping up: got any fun plans for the rest of the day? report on vocal fry in interactional contexts 7 i marked 3 vowels at the beginning and 3 at the end of each section. this added up to 78 vowels in each recording, or 2,496 total measurements. i generally marked the first or last three vowels, in order, but often skipped highly reduced vowels in favor of more clear vowels with a longer duration. i chose to include common filler and grammar words as long as they were not exceptionally reduced. most speakers used the same few words at the edges of utterances, such as “um,” “probably,” “yeah,” “so,” “that,” “i,” “like,” and “would.” in spite of this, i was able to collect a reasonable distribution of a variety of vowels, albeit with a higher degree of front vowels and slightly more high vowels. the vowel distribution for each speaker did not seem notably unequal. see figures 3-5. figure 3. vowel distribution 0 50 100 150 200 250 300 350 400 450 500 550 600 650 oɪ ɔ aʊ ʊ e u ɛ ɑ o i æ ɪ aɪ ə/ʌ catherine heidi dana tania landon andrew seth rob colorado research in linguistics, volume 25 (2021) 8 figure 4. fronting figure 5. height i ran a nasality script (styler, scarborough, johnstone, et al. 2014) on the annotated sound files in order to collect harmonic and formant amplitudes for comparison. the values i pulled from the script were the amplitude of harmonic 1 (h1), harmonic 2 (h2), formant 1 (a1), formant 3 (a3), and the fundamental frequency for pitch (f0). i subtracted h2, a1, and a3 from h1 respectively, expecting that more negative values would correlate with a higher degree of creak, and when it occurred with a low f0, may indicate vocal fry. i ended up focusing primarily on the h1-h2 value along with f0, but the h1-a1 and h1-a3 values served as verification of my methods, as they seemed to follow a similar pattern. i also divided each recording into three sections based on the time point of the first and final measurement of each speaker, and recorded whether the sound occurred early, mid, or late in the session. 3. results women had a lower mean h1-h2 value than men, at -8.409 (sd = 13.07) for women as compared to -4.716 (sd = 10.67) for men. with statistical significance being regarded as p < 0.05, this difference was found to be statistically significant at p < 0.001. women had a higher mean f0 value than men, at 196.305 (sd = 44.63) as compared to 126.835 (sd = 39.24). this was statistically significant at p < 0.001. thus, women showed more evidence of creak, but men displayed a lower front central back high mid low report on vocal fry in interactional contexts 1 pitch. although one would expect creak and low pitch to occur together in vocal fry, this is not surprising, as men have a biological precedent for lower pitch overall. see table 1 and figures 67. table 1. gender speaker h1-h2 f0 mean std. dev. mean std. dev. female -8.409 13.066 196.305 44.631 male -4.716 10.672 126.835 39.236 figure 6. h1-h2 by gender figure 7. pitch by gender younger speakers showed a slightly lower mean h1-h2 (-6.714, sd = 12.36) as compared to the older group (-6.427, sd = 11.78) and a lower mean f0 value (159.752 hz, sd = 56.23 vs 163.659 hz, sd = 52.72). with statistical significance being regarded as p < 0.05, the difference in h1-h2 between age groups was not found to be significant at p = 0.55. however, there was a statistical difference between groups in f0 at p = 0.02. thus, younger speakers can be said to have a lower pitch than older speakers, but there is no significant difference in creak. see table 2 and figures 8-9. -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 female male am pl itu de speaker gender 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 female male f0 (h z) speaker gender colorado research in linguistics, volume 25 (2021) 2 table 2. age speaker h1-h2 f0 mean std. dev. mean std. dev. older -6.427 11.780 163.659 52.715 younger -6.714 12.364 159.752 56.234 figure 8. h1-h2 by age figure 9. pitch by age young women had the lowest mean h1-h2 value (-10.530, sd = 14.84), and young men had the highest (-2.886, sd = 7.51). there was little difference between men (-6.55, sd = 12.84) and women (-6.306, sd = 10.63) in the older category, with men being slightly lower. with statistical significance being regarded as p < 0.05, differences between almost all values were found to be statistically significant at p < 0.001, except between older women and older men, in which case p = 0.71. thus, young females can be said to have a higher level of creak than older speakers and young males, and young males show less creak than either group, but there is no significant difference in creak between genders in the older group. gender again played a bigger role in f0 values, with men showing the lowest mean f0 values and women showing the highest. age was also consistent, with younger speakers showing the lower mean f0 value in their gender category. therefore, young men had the lowest value at 123.942 hz (sd = 38.34), then older men at 129.73 hz (sd = 39.93), younger women at 195.45 hz (sd = 47.93), and older women having the highest value at 197.16 hz (sd = 41.12). all differences crossing gender boundaries were found to be -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 older younger am pl itu de speaker age 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 older younger f0 (h z) speaker age report on vocal fry in interactional contexts 1 statistically significant at p < 0.001. the difference between young men and older men was also found to be significant at p = 0.003, but there was not found to be a significant difference between young women and older women (p = 0.53). thus, young men can be said to have the lowest pitch, followed by older men, and women have a higher pitch with no difference dependent on age. see table 3 and figures 10-11. table 3. gender and age speaker h1-h2 f0 mean std. dev. mean std. dev. older female -6.306 10.634 197.156 41.116 older male -6.550 12.844 129.732 39.932 younger female -10.530 14.836 195.448 47.933 younger male -2.886 7.506 123.942 38.342 figure 10. h1-h2 by age and gender figure 11. pitch by age and gender among older women, those speaking to those in the same age group showed the lowest mean h1h2 values. values were lower when speaking to women than to men. when partnered with an older female the mean h1-h2 value was the lowest at -8.031 (sd = 10.93), followed by older male partners at -6.536 (sd = 9.21), younger female partners at -5.589 (sd = 12.02), and younger male partners resulted in the highest mean at -5.114 (sd = 10.06). mean f0 values did not follow a -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 female male female male older younger am pl tu de speaker 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 female male female male older younger f0 (h z) speaker colorado research in linguistics, volume 25 (2021) 2 particular pattern, with older female partners resulting in the lowest mean f0 value at 193.526 hz (sd = 44.09), followed by younger male partners at 194.099 hz (sd = 44.03), younger female partners at 198.833 hz (sd = 39.26), and the highest mean f0 value occurring with older male partners at 202.282 hz (sd = 36.18). with statistical significance being regarded as p < 0.05, differences between partner demographics were mostly insignificant. the only statistically significant differences occurred when comparing h1-h2 of older female and young male partners, for which p = 0.02, and when comparing f0 of older female and older male partners, for which p = 0.03. thus, it can be said that partner demographic does not affect vocal fry in older women, except that there is more creak when the speaker is paired with the same demographic (older women) as compared to the opposite demographic (younger men), and gender seems to have an effect on pitch among same-age partners, causing lower pitch when speaking with older women than with older men. see tables 4-5 and figures 12-13. table 4. older females partner h1-h2 f0 mean std. dev. mean std. dev. older female -8.031 10.926 193.526 44.091 older male -6.536 9.210 202.282 36.180 younger female -5.589 12.020 198.833 39.255 younger male -5.114 10.064 194.099 44.032 table 5. paired two-tailed t-test comparing conversational partners among older women speakers older female vs young male older female vs young female older female vs older male young female vs older male young female vs young male older male vs young male h1-h2 p = 0.07 p = 0.02 p = 0.16 p = 0.41 p = 0.72 p = 0.17 f0 p = 0.21 p = 0.98 p = 0.03 p = 0.36 p = 0.27 p = 0.07 report on vocal fry in interactional contexts 1 figure 12. h1-h2 in older women figure 13. pitch in older women among younger women, there was no particular pattern in h1-h2 values. when partnered with an older female the mean h1-h2 value was the lowest at -11.869 (sd = 14.64), followed by younger male partners at -10.325 (sd = 14.67), younger female partners at -10.263 (sd = 14.69), and older male partners resulted in the highest mean at -9.664 (sd = 15.385). mean f0 values were lowest when speaking to another female, especially a younger female, and highest with a male, especially an older male. young female partners resulted in the lowest mean f0 value at 191.255 hz (sd = 49.42), followed by older female partners at 193.25 hz (sd = 45.07), young male partners at 195.942 hz (sd = 46.94), and the highest mean f0 value occurring with older male partners at 201.372 hz (sd = 49.98). with statistical significance being regarded as p < 0.05, differences between partner demographics were mostly insignificant. the only statistically significant differences occurred when comparing f0 between young female and older male partners, for which p = 0.047. thus, it can be said that partner demographic does not affect vocal fry in young women, except that they use a lower pitch with the same demographic (young women) as compared with the opposite demographic (older men). see tables 6-7 and figures 14-15. -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 female male female male older younger am pl itu de conversational partner 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 female male female male older younger f0 (h z) conversational partner colorado research in linguistics, volume 25 (2021) 2 table 6. younger females partner h1-h2 f0 mean std. dev. mean std. dev. older female -11.869 14.643 193.250 45.071 older male -9.664 15.385 201.372 49.982 younger female -10.263 14.694 191.255 49.424 younger male -10.325 14.665 195.942 46.938 table 7. paired two-tailed t-test comparing conversational partners among young women speakers older female vs young male older female vs young female older female vs older male young female vs older male young female vs young male older male vs young male h1-h2 p = 0.29 p = 0.26 p = 0.09 p = 0.59 p = 0.98 p = 0.63 f0 p = 0.65 p = 0.59 p = 0.09 p = 0.047 p = 0.39 p = 0.34 figure 14. h1-h2 in younger women figure 15. pitch in younger women among younger men, h1-h2 values were lowest when speaking with another male. when speaking to a male, h1-h2 was lower with young men than older men, and when speaking to a female, it was lower with older women than younger women. when partnered with a young male -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 female male female male older younger am pl itu de conversational partner 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 female male female male older younger f0 (h z) conversational partner report on vocal fry in interactional contexts 1 the mean h1-h2 value was the lowest at -3.872 (sd = 7.4), followed by older male partners at 2.813 (sd = 6.35), older female partners at -2.752 (sd = 8.03), and young female partners resulted in the highest mean at -2.101 (sd = 8.08). mean f0 values were lowest when speaking to a young partner, and highest with older partners. young male partners resulted in a lower f0 value than young female partners, and older female partners resulted in a lower f0 value than older male partners. young male partners resulted in the lowest mean f0 value at 118.058 hz (sd = 34.24), followed by young female partners at 122.665 hz (sd = 39.32), older female partners at 125.244 hz (sd = 35.61), and the highest mean f0 value occurring with older male partners at 129.795 hz (sd = 43.02). with statistical significance being regarded as p < 0.05, differences between partner demographics were mostly insignificant. the only statistically significant differences occurred when comparing f0 between older female and young male partners, for which p = 0.03, and between young males and older males, for which p = 0.001, and when comparing h1-h2 between young males and young females, for which p = 0.047. thus, it can be said that partner demographic only affects vocal fry for young men when paired with the same demographic (young men), which produces more creak, as compared to young women, and pitch is lowered when partnered with the same demographic as compared to older partners. see tables 8-9 and figures 15-16. table 8. young males partner h1-h2 f0 mean std. dev. mean std. dev. older female -2.752 8.025 125.244 35.609 older male -2.813 6.354 129.795 43.021 younger female -2.101 8.081 122.665 39.317 younger male -3.872 7.403 118.058 34.237 table 9. paired two-tailed t-test comparing conversational partners among young male speakers older female vs young male older female vs young female older female vs older male young female vs older male young female vs young male older male vs young male h1-h2 p = 0.39 p = 0.23 p = 0.94 p = 0.39 p = 0.05 p = 0.15 f0 p = 0.52 p = 0.03 p = 0.22 p = 0.10 p = 0.27 p = 0.01 report on vocal fry in interactional contexts 1 figure 15. h1-h2 in young men figure 16. pitch in young men among older men, there was no particular pattern in h1-h2 values. when partnered with a young male the mean h1-h2 value was the lowest at -7.123 (sd = 12.56), followed by older female partners at -6.598 (sd = 13.51), older male partners at -6.421 (sd = 12.34), and young female partners resulted in the highest mean at -6.05 (sd = 13.04). mean f0 values were lowest when speaking to a female partner, and lower among young partners than older partners. young female partners resulted in the lowest mean f0 value at 121.994 hz (sd = 43.34), followed by older female partners at 128.936 hz (sd = 40.45), young male partners at 130.06 hz (sd = 40.02), and the highest mean f0 value occurring with older male partners at 137.833 hz (sd = 34.21). with statistical significance being regarded as p < 0.05, h1-h2 differences between partner demographics were all insignificant. however, most f0 comparisons did result in statistically significant results. older men used a significantly lower pitch with older females than with older males (p = 0.01), with young females as compared to older (p < 0.001) or younger (p = 0.03) males, and with young males as compared to older males (p = 0.02). thus, it can be said that partner demographics do not affect creak in older men, but there is a significant effect on pitch. older men use the lowest pitch when speaking to women as compared to men, and the highest pitch when speaking to the same demographic (older men) as compared to older women or young men. see tables 10-11 and figures 17-18. -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 female male female male older younger am pl itu de conversational partner 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 female male female male older younger f0 (h z) conversational partner colorado research in linguistics, volume 25 (2021) 2 table 10. older males partner h1-h2 f0 mean std. dev. mean std. dev. older female -6.598 13.511 128.936 40.453 older male -6.421 12.342 137.833 34.209 younger female -6.050 13.036 121.994 43.337 younger male -7.123 12.560 130.064 40.021 table 11. paired two-tailed t-test comparing conversational partners among older male speakers older female vs young male older female vs young female older female vs older male young female vs older male young female vs young male older male vs young male h1-h2 p = 0.56 p = 0.61 p = 0.85 p = 0.76 p = 0.25 p = 0.48 f0 p = 0.08 p = 0.71 p = 0.01 p < 0.001 p = 0.03 p = 0.02 figure 17. h1-h2 in older men figure 18. pitch in older men i also measured the effect of a sound’s position in an utterance on h1-h2 and f0. with statistical significance being regarded as p < 0.05, an early position has a significantly lower (p < 0.001) -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 female male female male older younger am pl itu de conversational partner 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 female male female male older younger f0 (h z) conversational partner report on vocal fry in interactional contexts 1 mean h1-h2 value at -8.082 hz (sd = 13.19), as compared to a late position, measuring at -5.058 hz (sd = 10.63). a late position, however, has a significantly lower (p < 0.001) mean f0 value, at 151.832 hz (sd = 53.62), as compared to an early position, measuring at 171.586 hz (sd = 53.64). thus, it can be said that an early position in an utterance indicates more creak, but higher pitch, whereas a late position in an utterance indicates less creak but a lower pitch. see table 12 and figures 19-20. table 12. utterance position position h1-h2 f0 mean std. dev. mean std. dev. early -8.082 13.191 171.586 53.642 late -5.058 10.634 151.832 53.623 figure 19. h1-h2 by utterance position figure 20. pitch by utterance position i also measured the effect of a sound’s position in a recorded session on h1-h2 and f0. with statistical significance being regarded as p < 0.05, an early position has a significantly lower (p < 0.001) mean h1-h2 value at -7.287 hz (sd = 11.47), as compared to a late position, measuring at -5.376 hz (sd = 12.47). there is not a significant difference in h1-h2 between an early and mid -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 early late am pl itu de position 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 early late colorado research in linguistics, volume 25 (2021) 2 position (p = 0.62), but a mid position has a significantly lower (p = 0.005) mean h1-h2 value at -6.984 (sd = 12.283) as compared to a late position. there is also a significant difference in f0 values dependent on session position, with an early position having a significantly lower (p = 0.01) mean f0 value at 160.106 (sd = 53.89) as compared to a late position, measuring at 164.294 (sd = 56.05). there is not a significant difference in f0 between an early and mid position (p = 0.15), but a mid position has a significantly lower (p < 0.001) mean f0 value than a late position. thus, it can be said that creak decreases and pitch increases over the course of a recorded session. see table 13 and figures 21-22. table 13. session position position h1-h2 f0 mean std. dev. mean std. dev. early -7.287 11.468 160.106 53.892 mid -6.984 12.283 160.874 53.580 late -5.376 12.469 164.294 56.048 figure 21. h1-h2 by session position figure 22. pitch by session position when measuring the difference in h1-h2 according to vowel quality, mean values decrease with backing and raising of the vowel. back vowels have the lowest mean h1-h2 value when -12.000 -11.000 -10.000 -9.000 -8.000 -7.000 -6.000 -5.000 -4.000 -3.000 -2.000 -1.000 0.000 early mid late am pl itu de position 0.000 25.000 50.000 75.000 100.000 125.000 150.000 175.000 200.000 225.000 early mid late f0 (h z) position report on vocal fry in interactional contexts 1 comparing vowel fronting at -7.749 (sd = 11.04), followed by central vowels at -6.918 (sd = 12.52) and the highest value occurring in front vowels at -6.04 (sd = 11.76). high vowels have the lowest mean h1-h2 value when comparing vowel height at -7.454 (sd = 11.69), followed by mid vowels at -6.925 (sd = 12.53) and the highest value occurring in low vowels at -4.612 (sd = 10.92). when measuring f0 values according to vowel fronting, there was not a particular pattern obvious in the data. central vowels had the lowest mean f0 at 124.551 hz (sd = 35.942), followed by front vowels at 125.584 hz (sd = 37.68), and back vowels having the highest f0 values at 129.952 hz (sd = 45.38). when comparing height, mean f0 values lowered with lowering of the vowel. low vowels had the lowest mean f0 value at 123.588 hz (sd = 40.55), followed by mid vowels at 124.166 (sd = 35.89), and high vowels at 126.512 (sd = 41.42). when regarding statistical significance as p < 0.05, there was not found to be any significant difference in f0 changes across vowel quality. however, the difference in h1-h2 between back vowels and front vowels was found to be significant at p = 0.02, as well as between mid vowels and back vowels at p = 0.001 and between high vowels and low vowels at p < 0.001. thus, it can be said that while vowel quality has no effect on pitch, creak increases in back vowels as compared to front vowels, and significantly increases as a vowel is raised. see tables 14-16 and figures 23-26. table 14. vowel measurements by fronting fronting h1-h2 mean std. dev. f0 mean std. dev. front -6.040 11.756 125.584 37.681 central -6.918 12.527 124.551 35.942 back -7.749 11.035 129.952 45.376 table 15. vowel measurements by height height h1-h2 mean std. dev. f0 mean std. dev. high -7.454 11.685 126.512 41.416 mid -6.925 12.529 124.166 35.891 low -4.612 10.920 123.588 40.546 report on vocal fry in interactional contexts 1 figure 23. h1-h2 by vowel fronting figure 24. f0 by vowel fronting figure 25. h1-h2 by vowel height figure 26. f0 by vowel height -12 -11 -10 -9 -8 -7 -6 -5 -4 -3 -2 -1 0 front central back am pl itu de vowel fronting 0 25 50 75 100 125 150 175 200 225 front central back am pl itu de vowel fronting -12 -11 -10 -9 -8 -7 -6 -5 -4 -3 -2 -1 0 high mid low am pl itu de vowel height 0 25 50 75 100 125 150 175 200 225 high mid low f0 (h z) vowel height colorado research in linguistics, volume 25 (2021) 2 table 16. paired two-tailed t-test comparing h1-h2 and f0 between vowel qualities front vs central central vs back front vs back high vs mid mid vs low high vs low h1-h2 p = 0.26 p = 0.26 p = 0.02 p = 0.46 p < 0.001 p < 0.001 f0 p = 0.74 p = 0.14 p = 0.23 p = 0.44 p = 0.86 p = 0.40 4. discussion i hypothesized that young speakers would use more vocal fry overall, and while they used significantly lower pitch than older speakers, the harmonics measure did not show a difference between the two groups, so i cannot confirm my first hypothesis and say that the pitch change was due to vocal fry. comparing gender differences in speakers, women used more creak than men, and men had a lower pitch overall. i hypothesized that young people of both genders would use vocal fry more than older people of both genders, and that among young people, women would use more, while among older people, men would use it more. my hypothesis that young women would use the most vocal fry was correct, but that was the only part of my second hypothesis that was correct, as young men actually used the least vocal fry, so age was not a good indicator. there was no significant difference between genders in the older group. young women used significantly more vocal fry than all older people, and older people used more than young men. while unexpected, this is a very interesting finding, suggesting that gender has recently emerged as a predictor of creak for younger generations. perhaps young women have increased their level of creak in comparison to the older baseline, and young men have decreased theirs in order to distinguish themselves from their young female peers (or vice versa). young women used a lower pitch than older women, and young men a lower pitch than older men, which fits the pattern of increased creak in younger groups, while accounting for biological differences. overall, possibly due to my small sample size of speakers, many of my findings related to age and gender in an interactional context did not point to the demographic of an interactional partner having an effect on the speaker’s use of vocal fry. most differences due to the demographic of a conversational partner were statistically insignificant, but those that were significant stood out as potentially being part of a pattern. older women used significantly more creak when speaking to report on vocal fry in interactional contexts 3 another older woman as compared to with a younger man, and they lowered their pitch when speaking to another older woman as compared to an older man. young women lowered their pitch with other young women as compared to with older men. young men lowered their pitch when speaking to other young men as compared to older women or men, and used more creak when speaking to another young man as compared to with a young woman. older men experienced a lot of interaction between pitch and partner demographic, using the lowest pitch with young women, and the highest pitch with other older men. i hypothesized that differences in use of vocal fry due to the partner demographic would be mostly consistent regardless of the speaker demographic, with all speakers using more vocal fry with young people, and all older people using it more with men while all young people used it more with women. while my findings were overall not overwhelming and lacking in statistical significance, probably due to a small sample size, the small differences that i did find point to an interesting interaction that is more dependent on the speaker demographic than the partner – opposite to my hypothesis. most groups used the most creak or lowest pitch with the same demographic, and the least with the opposite demographic. older women used the most creak and lowest pitch with other older women, and the least creak with young men. young women used the lowest pitch with other older women, and the least creak and highest pitch with older men. young men used the most creak and lowest pitch with other young men. older men did not quite fit this pattern because of their unique pitch interactions, but they did use the least creak with young women. older men had the opposite effect in pitch, using the highest pitch with other older men and the lowest with young women. other interesting interactions occurred in the data aside from those related to speaker and partner differences. i hypothesized that measurements taken at the end of an utterance would have more vocal fry than those at the beginning. strangely, i found the opposite pattern in h1-h2 differences, showing significantly more creak on early words, yet a much lower pitch on late words. i would expect vocal fry to occur phrase-finally, and for low pitch and greater creak to correlate. i did take h1-a1 and h1-a3 measurements as well (calculating the difference between the first harmonic and first formant or third formant), which should also indicate creak, and while i did not run an analysis on these figures at this time, a brief glance at the results seemed to indicate that they generally followed the same pattern as h1-h2 values. an idea i had was that because i included filler words in my measurements, perhaps i was picking up a disproportionate amount of colorado research in linguistics, volume 25 (2021) 4 creak from the initial “um” that most people produced at the beginning of an utterance. however, after removing schwa measurements from the data, my results seemed to follow largely the same pattern, at least in early and late values. i am not sure what else would contribute to this incongruency in pitch and creak, but the numbers are very significant. i also hypothesized that position in an overall conversation would have an effect on vocal fry as well. i expected vocal fry to increase over the course of a conversation, due to speakers warming up to one another and relaxing their voices. in fact, i again actually found the opposite result, and both creak and low pitch significantly decreased over the course of a conversation. finally, i measured the difference in vocal fry due to vowel quality, expecting to find little difference between vowels, except higher levels of fry on schwa (due to initial “um”). i found that vocal fry increases as a vowel becomes more high and more back, so /u/, /ʊ/, and /o/ would have the greatest amount of creak. this is surprising to me, as i expected more fry on common words people used while thinking, such as “um,” “probably,” and “yeah.” the examples i can think of would be “so,” and “would,” occurring in contexts like “so, yeah” and “i think i would…”, but i am surprised that those would be the most creaky if that is the case. while most of my hypotheses were not confirmed, i am intrigued by the patterns i have found here and am interested in doing more research on the topic. due to the scope of my data, this experiment was limited in many ways. first, eight participants representing four demographics is likely far too small a data pool to get conclusive results, so further research would need to be done with much larger groups of participants. participants were also not particularly diverse, being almost entirely caucasian, and while everyone was from different regions of the us, individual dialect influences seemed somewhat neutralized by a colorado accent in most speakers. diversity of sexual orientation could also help strengthen the data, as most of my participants were heterosexual, and where they were not, my experiment design did not account for differences in vocal quality due to socialized norms of other sexual orientations or any expressions of gender outside the norm or the binary. also, being but one person doing annotations and measurements, 78 vowels from each of the 32 recordings was the extent of what time permitted me to sample for this initial study, but considering the recordings are 15-30 minutes long, there is a vast amount of data left that has not been measured. in the future, i would like to have these recordings completely annotated and measured for a deeper foray into this set of data. i would also like to examine “middle” measurements in addition to those at the beginning and end of utterances. in those report on vocal fry in interactional contexts 5 samples i did take, my vowel selection process could have been more consistent, as i did not always choose exactly the first and last three vowels to measure, but the “clearest” vowels that were closest to those. i often skipped vowels that were particularly reduced or in some way unclear. i also noticed that almost all the words i measured included common filler words such as “um”, “probably”, “so”, and “yeah”, so measurements of a greater variety of words with greater semantic content could likely improve this study as well. i was also not confident in all of the vowel markings i made, as some distinctions were hard to make between vowel and consonant when the vowel was followed by approximants with strong formant patterns. in many cases, i relied on marking by ear rather than by the spectrogram, but i would expect that some vowel measurements may have included consonantal content. there were a few inconsistencies in my process. i tried to only pair people that had not met before, but i had to make two exceptions due to problems with scheduling meetings. heidi is the mother-in-law of tania’s sister, and while they were friendly and familiar, i felt that they were not so familiar that they would not adhere to normal vocal patterns they used with others. catherine and rob, however, are romantic partners, but i could not find a way to avoid this pairing. they noticeably spoke to one another quite differently than they had with everyone else, so i don’t think this set of data was particularly strong in the study. there were other minor inconsistencies in the data collection, such as speakers using a different device to log into a session, a camera not working in one session, and my partner supervising a couple meetings while i was unavailable. i felt that what i saw and heard in the data did not necessarily always match the results. i felt, for instance, that andrew and landon had strong final vocal fry, possibly more than the older women group and certainly more than the older men group, but the results show them as the group with the least overall fry. i also felt that the fact that there was more creak measured on early words in an utterance was extremely strange, as i certainly saw and heard very strong creak on phrasefinal words in many of the speakers and not on early words, but no speaker showed a greater level of creak at the end according to the data. tania, for instance, seemed to have extremely strong vocal fry at the end of her utterances, and i heard her initial utterances as very high pitched and modal, but her results pointed to a particularly strong early creak and much lower final creak. this leads me to wonder if varying methods of measuring creak, or measuring for different types of creak as defined by keating et al., may have been beneficial for these particular speakers. colorado research in linguistics, volume 25 (2021) 6 vocal fry is a phenomenon that is largely misunderstood and misrepresented. additional research will help to build understanding and acceptance of it as a legitimate linguistic tool, and i hope that this study has done its part to take one small step in that direction. references anderson, rindy & casey klofstad. 2014. vocal fry may undermine the success of young women in the labor market. plos one 9.5. doi:10.1371/journal.pone.0097506 callier, patrick & robert podesva. 2015. multiple realizations of creaky voice: evidence for phonetic and sociolinguistic change in phonation. new ways of analyzing variation 44. online: https://stanford.edu/~podesva/documents/nwav2015-1up.pdf chao, monika, & julia bursten. 2021. girl talk: understanding negative reactions to female vocal fry. hypatia 36(1). 42-59. doi:10.1017/hyp.2020.55 esposito, christina. 2010. variation in contrastive phonation in santa ana del valle zapotec. journal of the international phonetic association 40(2). 181-198. doi:10.1017/s0025100310000046 keating, patricia, marc garellek, & jody kreiman. 2015. acoustic properties of different kinds of creaky voice. proceedings of the 18th international congress of phonetic sciences. online: http://idiom.ucsd.edu/~mgarellek/files/keating_etal_2015_icphs.pdf garfield, bob. 2013. “old fart” responds to the great vocal fry outcry of 2013. slate. online: https://slate.com/human-interest/2013/01/young-women-and-vocal-fry-slate-podcastwars-continue.html gordon, matthew, & peter ladefoged. 2001. phonation types: a cross-linguistic overview. journal of phonetics 29. 386-406. online: http://gordon.faculty.linguistics.ucsb.edu/phonation.pdf gross, terry. (host). 2015. fresh air weekend: weighing in on what it means to ‘sound gay’. [radio broadcast episode]. online: https://www.npr.org/2015/07/11/421470117/fresh-airweekend-weighing-in-on-what-it-means-to-sound-gay higdon, michael. 2016. oral advocacy and vocal fry: the unseemly, sexist side of nonverbal persuasion. legal communications and rhetoric. jalwd 13. 209-220. online: https://heinonline.org/hol/p?h=hein.journals/jalwd13&i=212 report on vocal fry in interactional contexts 7 irons, sarah & jessica alexander. 2016. vocal fry in realistic speech: acoustic characteristics and perceptions of vocal fry in spontaneously produced and read speech. the journal of the acoustical society of america 140.3397. doi:https://doi.org/10.1121/1.4970891 reynolds, eileen. 2015. what’s the big deal about vocal fry? an nyu linguist weighs in. new york university. online: https://www.nyu.edu/about/newspublications/news/2015/september/lisa-davidson-on-vocal-fry.html styler, will, rebecca scarborough, sarah johnstone, et al. 2014. automated nasality measurement script package. cu phonetics lab. wolk, lesley & nassima abdelli-beruh. 2011. habitual use of vocal fry in young adult female speakers. journal of voice 26.3. e111-e116. doi:https://doi.org/10.1016/j.jvoice.2011.04.007 1. introduction this paper investigates the relationship between language, culture, and cognition via metaphoric conceptualizations in quechua. a conceptual metaphor is a set of correspondences between a source and target domain used to reason about abstract concepts. cognitive linguists theorize that conceptual metaphors are not simply linguistic artifacts. instead, conceptual metaphor is part of the human conceptual system. the potential universality of some conceptual metaphors is thought to be reflective of the shared human experience rather than linguistic similarity (lakoff and johnson 1980, kövecses 2002). however, fernandez (1991) argues that cognitive linguists overemphasize the basic conceptual nature of metaphors, noting that claims of universality are based on a few examples from a handful of primarily indo-european languages. moreover, cases disproving universally shared conceptualizations are often overlooked. kövecses (2010a) questioned this disconnect, investigating the relationship between metaphor and culture by asking questions such as “which metaphors are universal and why?” and “what are the main dimensions along which metaphors vary?” claims of universality are predicated on the assumption that all cultures, or communities that share a set of practices and customs, including language, rely on figurative language, with no real consideration of the degree to which such an assumption may be true. figurative language is common in indo-european languages, the most widely studied language family. for example, in english, roughly one out of every twenty words in written text is metaphoric, with frequency estimates rising to nearly one out of every five words in casual, spoken discourse (pollio et al. 1990). native english speakers are largely unaware of the prevalence of figurative language and do not question whether their reliance on figurative language is culturally-based, assuming all cultural groups use figurative language with a similar frequency. however, the value placed on abstract versus concrete language is culture specific. for example, quechua places a much higher value on concreteness than abstraction. nature, which is central to everyday life, plays an analogic and symbolic role. however, while analogies to nature are made and may even become lexicalized in the form of a proverb or idiom1, conceptual metaphors are not as common in everyday quechua speech. this paper will identify two major themes of metaphor in southern conchucos quechua2: the conceptualization of time and space as a single, interconnected entity and the role of nature as a source domain for events in everyday life. a brief mention will be made of expected “universal” metaphors that seem to be “missing”, leading to a discussion of the greater role played by culture when mythopoetic concreteness is valued more than abstractness. 2. background few researchers have questioned claims that certain conceptual metaphors are found in all human languages (cf. gevaert 2005, kövecses 1990, 2000, 2005, 2010a, 2010b). while an investigation of universality claims would be inherently difficult as it would require a comparison between all languages of the world, the assumption of language-indifferent, shared conceptualizations seems logical since humans, regardless of language or culture, have a shared physiological experience. for example, when one becomes angry, physiological changes take place such as an increase in blood pressure, a feeling of increased internal temperature, possible visible physiological changes, such as the reddening of the face, and even an increase in preand post-perceptual executive functions such as a subconscious fight or flight response or an increased ability to consciously control attention (gazzaniga et al. 2002, kövecses 2000) kövecses (2000, 2005, 2010b), emphasizes the role that anthropological investigation should play in claims of conceptual universals realized through language. recognizing both the logic of a shared human experience and the role of culture in language use, kövecses urges caution in assumptions of universality, recognizing that the relationship between language and culture is something of a “chicken or the egg” question. kövecses (2000, 2005, 2010b) identified and investigated some seemingly universal metaphors, such as happiness is up, anger is a heated fluid, diseases are physical objects, and the body is a container. while these metaphors are found in germanic, romance, turkic, japonic, and sino-tibetan languages, conceptualizations differ slightly between cultures. for example, while the anger is a heated fluid and the body is a container metaphors are found in chinese and japanese, the belly serves as the center of the anger and the focus is on pressure build up more than heat. this differs from the english usage where the head, which releases pent up heat, serves as the container. conceptualization of the stomach as the center of anger in japanese is attributed to chi (kövecses 2005). by contrast, the focus on the temperature of anger found in european languages is attributed to the greek humors, a claim supported by diachronic use (geeraerts & grondelaers 1995). the influence of chi and humoral doctrine demonstrate the cross-cultural influences. gavaert (2001, 2005) took the role of culture one step further, questioning the universality of metaphoric conceptualization within a language. gavaert (2001, 2005) noted that latin and greek writings from 850-950 primarily utilized the anger is heat metaphor but written records from 950-1400 primarily used the anger is pressure metaphor. use of the anger is heat metaphor began to increase in the 1300s and, in the 1400s, regained its status as the more dominant of the two. surely, the human experience did not change between 850-1400, so the human experience itself is not sufficient to explain metaphor use. thus, while shared human experiences and limitations impact language, culture plays an equally important role in which aspects of the shared human experience are highlighted. 3. quechua metaphor survey to what extent does culture impact metaphor use in quechua? data for this project was elicited in two ways. the first was with the aid of a publicly available metaphor elicitation tool (gil and shen n.d.). the second method was through conversations on a number of topics not specifically related to figurative language that were later transcribed and searched for naturally occurring metaphors. data collection began with the metaphor elicitation tool. survey responses can be sent in for compilation in the creation of a world-wide study of figurative language use. the survey has two sections. in section one, consultants are asked to list as many terms as possible for a list of words. this list contains perceptual terms (ex. see, hear, listen, smell, taste, touch), sensory terms (ex. silent, noisy, light, dark, bitter, sweet, sour, hot, etc.), body part terms (ex. head, heart, eyes, brain, mouth, etc.), textural terms (ex. smooth, rough, bumpy, etc.), food terms (ex. eat, drink, digest, etc.) and travel terms (ex. crossroads, journey, dead-end, etc.). in the second section, consultants are asked to name as many metaphoric expressions as possible for three specific domains: emotion, mental states/activities, and time. examples of emotion metaphors in english include proposed universals such as explode with anger (anger is a heated substance in a container) or things are looking up (happy is up). examples of mental states/activities include have in mind and grasp the idea (ideas are physical objects). example of time metaphors include the festival is only two weeks away (time is motion) and invest a lot of time (time is a physical commodity). the following subsections report the most significant findings from the metaphor elicitation tool (gil and shin n.d.), discussing linguistic evidence of differing cultural outlooks. of note, only one survey of quechua using the metaphor elicitation tool has been published (owen 2020). however, this work focused on the cuzco dialect3. in order to add to a limited body of work, when applicable, comparisons will be made between findings presented here and those of owen (2020) as our findings differ in culturally and dialectally significant ways. 3.1. emotion according to kövecses (2005), the intense emotions are heat category of metaphors, which includes anger is a heated substance in a container and happiness is up, is among those likely to be universal. however, he notes that the image schema employed differs between cultural groups. for example, the heart serves as the center of emotion for most indo-european languages while the stomach generally serves as the center of emotion in asian languages. in both cuzco and southern conchucos quechua, the heart is conceptualized as the center of emotion. (1)-(3) demonstrate emotion metaphors found in cuzco quechua, as reported by owen (2020:6-8). in (1), a selfish person is conceptualized as having a hardened heart. example (2) demonstrates the existence of the potentially universal anger is a heated substance metaphor. in (3), the happiness is up metaphor is demonstrated. (1) kata sonko4 boulder heart “a selfish, hateful person” (2) sonko rawrari-shan heart burn-prog “an angry/upset person” (3) t’chosak sonko empty heart “a loveless person” southern conchucos quechua expands on the cuzco usage by lexicalizing a difference between the physical heart and emotional heart, as seen in (4a)-(4b). interestingly, even though a different word, chaki, exists for the emotional heart in southern conchucos quechua, the literal heart, shungo, can be metaphorically extended. however, unlike the cuzco dialect, which uses sonko in all literal and figurative contexts, shungo can be metaphorically extended only when it carries a positive connotation. this is demonstrated in (5a)-(5b). in (5a), shungo is used figuratively while (5b) evokes its expected literal meaning. this is not possible for negative value judgements. for example, in (6a) the emotional heart, chaki, can be conceptualized as the holder of hate. however, shungo, which can be used both literally and figuratively when serving as the container for positive emotions, cannot be extended as the container for negative emotions (7a)-(7b). (4a) chaki (4b) shungo heart heart “center of emotion” “human cardiac muscle” (5a) aliyen shungo (5b) aliyen shungo good heart good heart “be kind hearted” “the physical heart is in good condition” (6a) chikikumi “a person who hates” (7a) *alitsu shungonshi (7b) alitsu shungonshi *bad heart bad heart *“a heart that is hateful” “the physical heart is unwell” sadness is the most common theme in quechua songs, poetry, and proverbs. areal contact and generational differences do not seem to have impacted the comparisons and conceptualizations used to understand and convey sadness between the southern conchucos and cuzco dialects. in cuzco quechua, there is a slight tendency towards the use of sensation-related source domains, such as intense emotions are heat (8). while this metaphor is used in southern conchucos quechua, it is not common in everyday speech. instead, emotions are more often compared to nature. for example, (8)-(9), from lara (1969:226), show a common comparison between tears and bodies of water. (8) shows the comparison between tears and rain where the tears fall at an intensity similar to rain. in (9), intense sorrow is compared to a river that floods as a result of copious crying. example (10) is a proverbial idiom, first use in a song about sorrow (owen 2020:8). due to its idiomatic status, (10) conveys a meaning more intense than one may expect from the metaphoric mappings on their own, expressing severe sadness to the point of despondence. in addition to a comparison between a flood of tears and the level of water in a river, (10) indexes the divided self metaphor in which one’s heart is carried away in sorrow while one’s corporeal body remains. (8) para gina waqanaypay5 “to see me cry like the rain” (9) mayu gina in waqarqan “like a river in flood i sob” (10) wekaki majun apaj-uwa-ʃan “the river of tears is taking me away” 3.2. cognitive events while not specifically identified as a likely universal, the ideas are physical objects metaphor is very common cross-linguistically and is often discussed as if its universality is a forgone conclusion. this metaphor is found in cuzco quechua, as reported by owen (2020:10) in (11) below. (11) erkeka manam hapin yatashe-ta “the child did not grasp the idea” interestingly, while (11) can be translated into southern conchucos quechua, the phrase would never be used as its meaning is culturally anathema. a phrase such as grasp the idea entails an implicit agreement to a shared cultural judgement regarding new or creative ideas. this type of phrase conveys a positive value judgement associated with the learning of that which is new or creative. however, new ideas are not culturally valued. instead, they represent a rejection of one’s past, one’s culture, and one’s heritage. after all, if an idea is culturally preserved, why do you need to introduce something new? as our consultant noted, "ideas are new and creative. that is not part of the cultural heritage passed down.” (loayza 2021). she explained that the cuzco dialect is spoken by younger generations who are more open to outside ideas, thus accounting for its presence in owen (2020). 3.3. time perhaps the most widely accepted metaphoric universal is that of time as space (lakoff & johnson 1980, lakoff 1993). it has long been claimed that all languages rely on similar conceptualizations of time and space, with the future situated in front and the past behind, even if the language primarily relies on a vertical image schema (lakoff & johnson 1980, lai & boroditsky 2013). for example, chinese time/space metaphors offer surface evidence against claims of universality because the primary conceptualization of time arguably relies on spatial verticality. however, further study has shown that speakers shift between both a vertical and horizontal scale during initial stages of cognitive processing, demonstrating that both directional and vertical scales constitute primary conceptual categories (lai & boroditsky 2013). more serious challenges to universality assumptions come from languages such as aymara. while aymara uses the expected horizontal scale when discussing time, the future is conceptualized as behind while the past is in front (núñez 2003, núñez & sweetser 2001, 2006). descriptions of time as space metaphors in quechua are contradictory. scholars such as almedia (2005), almedia & haider (2012), and estermann (1998) have claimed that time conceptualization in quechua follows a pattern similar to that found in aymara. these claims are based on the use of quechua words such as nawpa and qhipa. literally translated, nawpa means “front” while qhipa means “behind”. variations of the following sentences, taken from faller and cuéllar (2003:4), have been translated in support of such claims, receiving glosses such as in (17)-(18). (17) nawpa-q-qa allin-si kas-sqa front-1-top good-ev be-pst “the past was better” (18) qhipa-man-qa allin-si ka-n-qa rear-poss-top good-ev be-3-fut “the future will be better” however, such claims have been challenged by many, including faller and cuéllar (2003) and owen (2020), who claim that no evidence of this reversed directionality is found in everyday speech. interestingly, while owen (2020) states that reverse directionality is not found in everyday speech, she notes the existence of a few idiomatic phrases that exhibit the conceptualization of the future as behind and the past as in front, and questions whether their existence might serve as evidence of a dual conceptualization of time. however, discussion with our consultant revealed that these idioms which seem to demonstrate reverse directionality are noncompositional. noncompositional idioms are memorized, fixed expressions that people know as a whole, such as the english phrases by and large or kick the bucket in which meaning is not contributed by the individual words. therefore, such phrases cannot be used as evidence for an existing dual time conceptualization, where the past may be either in front or behind one. faller and cuéllar (2003) explained the time as space metaphor in quechua as varying along four dimensions: the horizontal dimension, the vertical dimension, the entity in motion (ego moving versus time moving), and the cyclical nature of time. while faller and cuéllar’s (2003) discussion of a complex system of time conceptualization due to an inextricable link to three-dimensional space of the interconnectedness of conceptualizations of time and space is very thorough, our consultant provided a more straightforward answer. she explained that prior translations of sentences such as (17)-(18) are not entirely correct. in her translation, the initial temporal clause establishes a time point for comparison to the main clause. for example, in (19), nawpa can be translated as “in front/first/ahead of” and the tense marked on the main clause is past. however, the composite meaning is not one in which nawpa refers to “the past as in front”. instead, nawpa establishes that something that was said “first”, or “prior to another thing” which will be reported in the main clause. if such an example showed evidence of the conceptualization of the future as “in front”, one could claim that english shares this reverse directionality, as evidenced in a sentence such as (20)-(21), where the initial clause establishes the existence of something that happened ahead of the event in the main clause. (22) includes the translation of the english sentence in (21), with nawpa. finally, (23) shows our consultant’s translation of the qhipa example sentence, showing a similar process. the conceptualization of time in quechua deserves further investigation. while a more complex process may be at work, the more straightforward explanation provided by our consultant, which is reflective of native bilingual intuition, deserves equal consideration. (19) nawpa-q-qa allin-si kas-qa front-1-top good-ev said-pst “what i said first, they say it was good.” (20) those who lived first, lived well. (21) people lived well in the past, in the coming times, they will not live well. (22) nawpa-q-qa runa-kuna alli-n kawakuyaran, first-1-top person-pl good-poss live-pst, “in the past, people lived well, hamo-wata runakuna manam alli-n kawayanga future-year person-pl not good-poss live-fut in the future they will not live well” (23) qhipa-man-qa allin-si ka-n-qa last/next-poss-top good-ev say-3-fut “whatever comes after, they said it is going to be good” 4. the role of nature and geography the following subsections address the question of metaphors in quechua from a different direction. initial metaphor elicitation was aided by a metaphor elicitation tool for the purposes of identifying present and absent potentially universal conceptual metaphors (gil & shen n.d.). however, during conversations with our consultant, the pivotal role played by nature in both abstract and concrete language quickly became clear. while these topics are not included in the metaphor elicitation tool, excluding them from this study would result in an incomplete cultural portrayal as demonstrated through language use. the following subsections discuss metaphors that were freely produced over the course of multiple conversations not limited to metaphor elicitation. this methodology is important as it is a more natural reflection of native language use. while translations for elicited survey metaphors were possible for many proposed universals, such as intense emotions are heat, such source domains are not commonly used. instead, those most frequently used as well as those our consultant was most comfortable using shared the common theme of nature as a source domain. 4.1. life is a journey in english, we discuss the journey of life, often comparing life events to smooth or bumpy roads. southern conchucos quechua possesses a similar life is a journey metaphor. since the majority of roads in this region are not smooth, an easy or happy period of life is compared to a flat road covered in soft grass. while roads and paths are expected to be rocky, bumpy, and generally difficult to traverse, roads are particularly treacherous during the rainy season, when muddy. this has led to the establishment of a source domain of a muddy/sticky road that one must slog through which maps to difficult periods in one’s life. for example, a phrase such as (24), which translates roughly to “the muddy road i am stuck in”, can be used either literally or figuratively. the literal use refers to a physical path that is hard to traverse, resulting in one getting physically stuck. in the metaphoric instantiation, the literal path serves as the source domain, likening treacherous terrain to a difficult period of life, such as when one’s husband dies. (24) mitu, mitu n'yan-ni-pi muddy, muddy road-in-poss “the muddy road i am [stuck]6 in” 4.2. love is a journey like life, love is also conceptualized as a journey. unlike the journey of life, metaphors elicited for love seemed to focus on the journey’s trajectory or path, rather than the state of the path itself7. when asked about love, our consultant immediately produced three metaphoric phrases. in (26), a relationship is compared to a dead end. when one reaches the dead end, the relationship is over. this phrase must be used in context as it can also mean to literally die or to be figuratively dying (emotionally, but when not caused by the end of a relationship). in (27), relationship complications and confusion are compared to taking a road one should not have taken. confusion is also used as an idiomatic euphemism when referring to adultery. in (28), “confused lovers” refers specifically to those who fall in love while at least one partner is already in a relationship. (26) ushaka-kuumi to end/finsh-refl “i am dying/ending” (27) nanita pash-qa road be.confused-top “i took the road i should not have taken” (28) pantai walmiki confused lovers “confused lovers” 4.3. plant growth given the dependence on and respect paid to nature, one of the most strikingly absent metaphors is the comparison of the human life cycle to plant growth. many languages compare the observable, relatively short, cyclical, growth period of plants to intrapersonal and interpersonal development. for example, common english expressions include (29)-(31), each likening human development to the life cycle of plants. such a comparison is not possible in quechua. "a plant is a plant. a human is not a plant. a human does not grow like a flower." (loayza 2021). (29) our relationship blossomed. (30) he has shown a tremendous amount of growth this past school year. (31) however, do not leave old friendships unattended, for they will wither and die. in peru, flowers are highly regarded for their beauty. however, there is a strong distinction between the inherent respect that should be afforded gifts of nature and the use of nature as a symbol. while violent acts of nature, such as storms, serve as the source domain of many metaphors, gentle, delicate gifts of nature do not. for example, people, girls in particular, will pick flowers to put in their hair, hats, or clothes. the flower is “beautiful”, a trait culturally associated with the category “flower”. however, attributes of the flower are not bestowed upon the wearer – a girl wearing a flower is no more or less beautiful when adorned with this symbol of beauty. the restricted symbolic power of waitata (“flower”) is crucial as it explains the status of a flower as well as its preclusion from use as a symbol of beauty. flowers cannot be used in analogies or comparisons. while a boy may give a girl a pretty flower, she cannot be compared to the beauty of the gift. other aspects of flower, such as purity or delicacy, are also precluded from symbolic use. for instance, a particular yuraq waitata (“white flower”) is used to make a common medicine. the use of this flower is widely known so the flower is recognized for its healing properties. however, no aspect of the flower, from its color to the medicine it is used for, can be symbolically used to refer to cleansing or healing. similarly, towns, such as wari, are sometimes named after flowers specific to the area. wari was named for the waitata waganku, or “crying flower”. while the area is known for this flower, no comparison is made between sorrow, hardships, or other causes of sadness from waganku nor are any associations made between the town and waitata. 4.4. weather and geography weather plays a critical role in daily life, dictating whether life-sustaining plants will grow or die, whether rivers will swell or dry, and whether roads will be passable or treacherous. the import afforded weather is seen in its frequent use as a source domain. the complex relationship with the weather can be seen in (33)-(34), where the weather is alternatively conceptualized as a calming presence and compared to sorrow. chillap-yará, in (33), refers specifically to the sunrise, which brings a calming feeling of happiness. in (34), clouds are used to discuss difficulties and aimlessness, while tears are compared to falling raindrops. (33) chillap-yará happiness-bright (34) mamayri runnayawasqa my mother in the middle of the clouds para, p’uyu sunquyanpi and the rain had conceived me p’uyu gina muyunaypay to see me wander like the clouds para gina waqanaypay to see me cry like the rain (lara 1969:226) 4.5. geography the sun, moon, and mountains also play a symbolic role. while those descended from the inca are called “sons of the sun”, the passage of time is marked by the moon. at night, the sun “dies” (35). however, it is not reborn in the morning, instead, it “rises” (36). these phrases are also used to refer to the cardinal directions of “east” and “west”. (35) rupa wanu-n sun die-3.sg “west, where the sun dies (on a specific trajectory)” (36) rupa yuri-mu-n sun born-dir-3.pl “east, where the sun rises (on a specific trajectory)” similar to english, mountains are recognized for their majesty as well as the time involved in their development. the word for “mountain” is hilka. when found in combination with a person, such as in (37), it compares the age of a person to that of the mountain. the age of the mountain can also be used to refer to those who came before you, such as in (38). unne runa can be used without hilsa to refer to “people of the past” without necessarily meaning those in the distant past. unne runa can be used both literally and metaphorically. when une, “old”, is used in combination with hilsa, the meaning changes as seen in (39). (37) ruku runa, ruku hilsa old person, old mountain “old like the mountain” (38) unne hilsa runa old mountain people “people of the distant past” (39) une hilka old mountain “very old person” (literal) “wisdom of the mountain” (figurative) 4.6. animals animals play a significant role in quechua myths, serving as sources of symbolism in everyday language. the region in which southern conchucos quechua is spoken has fewer animals than the regions in which other dialects are spoken, resulting in a comparatively narrow range of animals in tales. foxes, owls, and donkeys traditionally carry the majority of symbolic weight, although condors, llamas, and guinea pigs can be found in some tales. in particular, while not native to this region, guinea pigs carry significant cultural value used by healers as they are thought to absorb bad energies and sickness. additionally, now-common domesticated animals such as the cat and dog are referenced in figurative language. the following sections will discuss the cultural associations encoded in phrases including the fox, donkey, and dog. the fox the fox is known as a sly trickster. while his slyness entails intelligence, the fox carries a negative connotation. for example, (40) includes the first five lines of a poem in which a man compares himself to a fox because he is hated by those around him. (40) ah! ato’k, ato’k, intj! fox, fox, “oh! fox, fox,” haliga ato, highland fox, “fox from the highlands,” gam-ta noga-ta runa chiki-man-si you-to comp-me man hate-obj-as “like you, i am hated by man” gam-ka chuki usha-n-ta umpaptiki suwap-ti you-gen hate sheep-poss-acc when steal-acc “they hate you when you steal their sheep” no-gata chiki wawanta suwap-ti comp-me hate daughter take-acc “and to me, they hate me when i steal man’s daughter.” the donkey the donkey is known for his stubbornness and general lack of intelligence. our consultant had no trouble immediately producing multiple idioms involving the donkey. the idiom in (43), which literally translates to “donkey’s ear”, is used to refer to one who is stubborn, unintelligent, or both. the phrase in (44) can be used either literally or figuratively. in literal context, it describes one who is klutzy. when used figuratively, it refers to one who makes poor life choices. in (45), a different aspect of the donkey is highlighted. while he is seen as stubborn and ignorant in general, he is also forced to work hard. when used within the domain of work, the negative connotation carried by the donkey is reversed. therefore, one who “works like a donkey” is not a stubborn worker but a hard worker. even though hard work is valued enough to override the negative attributes of a donkey, such as stubborn ignorance, comparison to a donkey is not particularly flattering and would refer to someone forced to do grunt work rather than someone in a position of power. finally, the idiom in (46) is a phrase our consultant remembers hearing from parents when they were particularly unhappy with their children. (46) is a demeaning phrase and is not commonly used unless one loses their temper entirely. while it translates roughly to “i will hit you so hard that you lose all of your intelligence, so hard you will not know who you are”, the intent is similar to the english phrase “slap you stupid” or “slap you into tomorrow”. (43) ashnu ringri donkey ear “[he is] stubborn and/or ignorant” (44) ashnu nupuli donkey walk “[he] walks like a donkey” (45) ashnu-ta-nu asu-si-ya-shu-ki-nashunki donkey nu-comp work-make-pl–sub-acc "[they are] made to work as if a donkey" (46) ashnu-man apta-si-shu-ki donkey-to send-make-sub-obj “[i will] send her to the donkey” the dog while not symbolic animals in a historical or traditional sense, both cats and dogs play a role in everyday figurative language, finding their way into modern proverbs and tales. one can be compared to a cat if they are “standoffish”. by contrast, a nosy person is compared to a dog. the phrase in (47) can be used literally, if a dog runs up to someone and sniffs them in inappropriate places, or figuratively, to refer to a nosy person who, like a dog, puts his nose where he should not, demonstrating the presence of the knowledge is a physical object and perception is reception metaphors. (47) alliqu muski-chi dog smell-na “dog sniffing where it should not” the “smelling dog” is a commonly used idiom. however, the figurative use of muski is not restricted to constructions including alliqu. instead, it can be used figuratively with humans, such as in (48), where the person is not compared to a nosey dog but is still figuratively sniffing where he should not. the -chi ending adds a negative value judgement for the action performed, not the scent itself. for example, muski-chi indicates that the smelling action is inappropriate, not that the scent itself is foul. the positive connotation marker -pa can be added to muski if the nosiness is warranted, such as a detective who “sniffs out” the perpetrator (49). -chi and -pa can be used only with animate agents, such as people or dogs who are considered equally volitional with respect to agency over controlling, or acting upon, their desire to sniff. (48) runa muski-chi he smelled-na “he sniffed it out” (49) muski-pa smell-pa “sniff out” 5. conclusion this paper has presented a survey of collected metaphoric and symbolic themes in quechua. it began with an investigation of claims of metaphoric universals, finding many expected universals while discussing the role of culture on metaphoric instantiation. the early discussion of expected metaphors progressed to an examination of the role of nature as well as the cultural implications of outside influences, further separating both the language and cultural outlook between southern conchucos and cuzco quechua. the symbolic role played by entities found in traditional myths, poetry, and songs, such as animals and nature make it difficult to define metaphoric language in quechua. what looks metaphoric to an english speaker may be built upon the concrete belief in a mythical world. what constitutes “concrete” in a mythopoetic model blurs the line between what is thought of as “metaphoric” in western cultures as it requires a suspension of western scientific reasoning with respect to what is considered to be fictive, abstract, or figurative (almeida & haidar 2012). unlike languages based in cartesian thought, which are based in “real science” and rely on conceptual metaphors, quechua values concreteness (almeida & haidar 2012). this is not to say that metaphor is not found. however, the idea of a traditional source domain may be misleading. in quechua, many conceptual mappings are more symbolic analogies linking everyday events to traditional, mythical beliefs. for example, one could conclude ashnu ringri (donkey ear) demonstrates the people are animals metaphor. however, animals, such as donkeys and foxes, are found in many myths and parables. in these tales, animals play a role interpreted by westerners as metaphoric. for example, the tale of the highland fox in (40) compares the feelings of a man to the mythical role of a fox, who is thematically portrayed as “sneaky” and “hated” in tales. while the fox serves as a symbol for “sneakiness”, this symbolism is based on a concrete and observable system of beliefs, rather than a complex, abstract mapping system linking relevant aspects of a source domain to a target domain based on underlying image schemas. entailed associations, such as the slyness of ato and the stubborn ignorance of ashnu, rely on shared knowledge of cultural teachings as much, if not more, than present-day observable occurrences, challenging more traditional interpretations of conceptual metaphor theory, in which conceptual metaphors are language-independent, resulting from shared image schemas. the quechua language challenges general assumptions of universality in figurative language. many metaphors commonly hypothesized to be universal are found in both southern conchucos and cuzco quechua, such as happiness is up and intense emotions are heated substances (if not necessarily the subcategory of anger is a heated substance). however, cultural value is placed on tradition, history, and mythical proverbs, blurring the line between what is conceptualized as a concrete link to a myth versus an abstract image schema relating the physical to that which is not tangible. thus, simple verification of the presence of proposed universal metaphors tells only a part of a larger, societal story. instead, we must also consider the role played by metaphor within cultural groups as the assumption that all cultures rely on figurative language does not seem to hold, at least for quechua but likely for many other languages as well. this area deserves further study on two fronts. first, how can we distinguish between that which is conceptualized as real or concrete from that which is metaphoric, without coloring our analysis with western scientific presumptions? second, what can differences in metaphor usage between southern conchucos and cuzco quechua tell us about cultural shifts and ideologies? references almeida, ileana. (2005). historia del pueblo kechua. quito: editorial abya yala. (2009). el modelo mito-poético del mundo en la cultura quechua durante la época del tawantin suyo. in: arcos cabrera, carlos (ed.), sociedad, cultura y literatura. 50 años flacso. quito: rispergraf c. a., 271–283. almeida, ileana, and julieta haidar. (2012). the mythopoetical model and logic of the concrete in quechua culture: cultural and transcultural translation problems. sign systems studies. 40(3/4), 484-513. https://doi.org/10.12697/sss.2012.3-4.12 alverson, hoyt. (1994). semantics and experience: universal metaphors of time in english, mandarin, hindi, and sesotho. baltimore: johns hopkins university press. boers, frank. (1999). when a bodily source domain becomes prominent. in r. gibbs, and g. steen (eds.), metaphor in cognitive linguistics. 47-56. amsterdam: john benjamins. estermann, josé. (1998). filosofía andina. quito: editorial abya-yala. faller, martina, and mario cuéllar. (2003). metáforas del tiempo en el quechua. in actas del iv congreso nacional de investigaciones lingüístico-filológicas. 1-11. gazzaniga, michael s.; richard b. ivry; and george, r. mangun. (2002). cognitive neuroscience: the biology of the mind. new york: norton. gil, david and yeshayahu shen. (n.d.). typological tools for field linguistics. retrieved from https://www.eva.mpg.de/lingua/tools-at-lingboard/questionnaire/figurativelanguage_description.php. accessed 2/11/2021. geeraerts, dirk and stefan grondelaers. (1995). looking back at anger. in: john r. taylor and robert e. maclaury (eds.), language and the cognitive construal of the world, 153179. berlin/new york: mouton de gruyter. gevaert, caroline. (2005). the anger is heat question: detecting cultural influence on the conceptualization of anger through diachronic corpus analysis. in n. delbacque, j. van der auwera, and g. geeraerts (eds.), perspectives on variation: sociolinguistic, historical, comparative, 195-208. berlin: mouton de gruyter. johnson, mark. 1987. the body in the mind. chicago: university of california press. kövecses, zoltán. (1990). emotion concepts. berlin and new york: springer-verlag. kövecses, zoltán. (2000). metaphor and emotion. new york and cambridge: cambridge university press. kövecses, zoltán. (2005). metaphor in culture: universality and variation. cambridge: cambridge university press. kövecses, zoltán. (2010a). variation in metaphor. ilha do desterro, 53: 13-39. kövecses, zoltán. (2010b). metaphor, language, and culture. delta: documentação de estudos em lingüística teórica e aplicada, 26(spe): 739-757. https://dx.doi.org/10.1590/s010244502010000300017 lakoff, george. (1993). the contemporary theory of metaphor. in a. ortony. (ed), metaphor and thought (2nd ed.). cambridge: cambridge university press, 202-251. lakoff, george and mark johnson. (1980). metaphors we live by. chicago: the university of chicago press. lakoff, george and mark johnson. (1999). philosophy in the flesh. new york: basic books. lakoff, george and zoltan kövecses. (1987). the cognitive model of anger inherent in american english. in d. holland and n. quinn (eds.) cultural models in language and thought. new york and cambridge: cambridge university press. lakoff, george and mark turner. (1989). more than cool reason. a field guide to poetic metaphor. chicago: university of chicago press. lai, vicky t., and lera boroditsky. (2013). the immediate and chronic influence of spatiotemporal metaphors on the mental representations of time in english, mandarin, and mandarin-english speakers. frontiers in psychology 4:142. doi: 10.3389/fpsyg.2013.00142 lara, jesus. (1969). la literatura de los quechuas. juventud. levinson, stephen c.; jürgen bohnemeyer; and n.j. enfield. (2001). 'time and space' questionnaire for 'space in thinking' subproject. in stephen c. levinson & n.j. enfield (eds.), manual for the field season 2001, 14-20. nijmegen: max planck institute for psycholinguistics. núñez, rafael. (2003). conceptual structures and cultural variation. metaphorical spatial construals of time in aymara. manuscript. núñez, rafael and eve sweetser. (2001). spatial embodiment of temporal metaphors in aymara: blending source-domain gesture with speech. proceedings of the 7th international cognitive linguistics conference. santa barbara, california. núñez, rafael and eve sweetser. (2006). aymara, where the future is behind you: convergent evidence from language and gesture in the crosslinguistic comparison of spatial realizations of time. cognitive science. 30, 410-450. owen, rosalind. (2020). sweet songs and soft hearts: metaphors in cuzco quechua. proceedings of the canada linguistics association. pérez, regina g. (2008). a cross-cultural analysis of heart metaphors. revista alicantina de estudios ingleses. 21:25–56. pedersen, duncan; hanna kienzler; and jeffrey gamarra. (2010). llaki and ñakary: idioms of distress and suffering among the highland quechua in the peruvian andes. culture, medicine, and psychiatry, 34(2): 279-300. saroli, a. (2005). the persistence of memory: traditional andean culture expressed in recurrent themes and images in quechua love songs. confluencia, 47-56. acknowledgments the project was conducted with the help of a native speaker of southern conchucos quechua, doris loayza, who is also a native speaker of spanish and a fluent english speaker. all southern conchucos examples were proved by doris, referred to as “our consultant”, who also provided invaluable insight into the pragmatic and sociocultural implications of using quechua or spanish. glossing conventions 1 – first person 3 – third person acc – accusative comp – comparative dir – direction ev – evidence intj – interjection fut – future tense gen – genitive na – negative action, animate entities only neg – negative obj object pa – positive action, animate entities only pl – plural poss – possessive prog – progressive pst – past tense sg – singular sub subject top – topic endnotes 1 an idiom is a phrase with a composite meaning that differs from what one would expect by combining the meaning that each word has outside of the phrase. unlike proposed universal conceptual metaphors, idioms are conventionalized phrases shared by members of a cultural group. 2 southern conchucos quechua is the primary dialect of the ancash region of peru. 3 cuzco quechua is the primary dialect of the cuzco region of peru 4 interlinear glosses provided by owen (2020) are not consistently glossed. the majority of examples included a phrase or sentence in quechua, a direct translation of each word, and a translation using english word order. when necessary, interlinear glosses with grammatical markings were included. due to differences in southern conchucos and cuzco quechua, examples credited to owen (2020) were included in their original published form. 5 lara (1969) includes poems in quechua. later translated into spanish by saroli (2005), no interlinear glosses have been agreed upon and therefore were not included. 6 this phrase is idiomatic and implies “being stuck” without including the word “stuck” in the phrase. 7 this may have been a coincidence due to elicitation of both life and love metaphors in the same session. our consultant may not have wanted to repeat that which she had already stated. 1. overview research on tagalog has shown that when describing transitive events, an event wherein an entity acts on another entity, speakers exhibit a strong preference for mapping undergoers (encompassing patients, themes, goals, etc.) instead of actors (agents, experiencers, causers, etc.) to the privileged syntactic argument function (as indicated by ang-marking). however, the role of individual verbs and their co-occurrence patterns with their actor and undergoer arguments within these voice structures remains an open question. to what extent do individual verbs prefer mapping undergoers to the privileged syntactic argument? to what extent does this preference vary across verbs and are potentially verb-specific behaviors modulated by referential properties known to affect undergoer and actor voice selection? this study uses corpora methods (gries & stefanowitsch 2004; bresnan et al. 2004; bybee 2006; colleman 2009) to examine the tagalog undergoer voice preference for frequently occurring semantically transitive verbs (e.g., bangga ‘bump,’ tawag ‘call,’ tulak ‘push,’ etc.). the data were extracted from the tltenten 2019 tagalog web corpus and coded for several morphosyntactic and semantic features. preliminary results (n = 10 verbs, tokens = 685) suggest that preference for ang-marked undergoers is not monolithic. each verb exhibits specific patterns of ang-marking undergoers and actors that vary somewhat per verb and the relative weighting of their arguments' referential features. furthermore, the contexts for mapping actors to the privileged syntactic argument appear to be much more highly constrained. these results suggest that complex interactive relationships between these factors (and others) must be examined in order to explain the undergoer and actor voice distributions in tagalog. 2. background & significance 2.1. tagalog background information tagalog is part of the central philippine subgroup of philippine languages and is part of the western-malayo-polynesian set of austronesian languages. it is native to manila, the largest city of the philippines and is, along with english, the lingua franca in many cities. tagalog is spoken by ~ 21.5 million speakers in the philippines (sauppe et al. 2013). as of 2008, it was estimated that over 90% of the population in the philippines is either a firstor second-language speaker of tagalog (schachter & reid 2008). speakers tend to be multilingual in tagalog, english, and/or another philippine language. 2.2. theoretical grounding and terminology this paper focuses on tagalog angand ng-/samarking1 on arguments, the co-indexation of those arguments on semantically transitive verbs via voice affixation, and the referential properties of those arguments. grammatical relations such as "subject" or "object" may not be applicable to tagalog (e.g., schachter 1976, 1977; schachter & otanes 1972; naylor 1995 kroeger 1993; himmelmann 2008). therefore, i draw on a few key concepts from the framework of role and reference grammar (rrg, e.g., foley and van valin, 1984, van valin and la polla, 1997, etc.). rrg defines two types of semantic roles: thematic relations in the traditional sense of agent, theme, patient, experiencer (fillmore 1968; gruber 1965) and generalized semantic roles called semantic macroroles. the macroroles play a central role by acting as the interface between the arguments in a verb’s structure (in rrg, logical structure) and syntactic representations. the two macroroles actor and undergoer each subsumes specific semantic relations. within rrg, grammatical relations such as subjects, objects, etc. are replaced with the notion of a privileged syntactic argument (psa), which is a “construction-specific relation and is defined as a restricted neutralization of semantic roles and pragmatic functions for syntactic purposes” (van valin 2002, p. 18). although rrg focuses on mapping of semantic roles to grammatical relations, patterns of semantic features, such as definiteness, animacy, topicality, etc., that more are often associated with these (macro)roles also play a significant role in how semantic roles are mapped to syntactic roles. with respect to tagalog, ang-marked arguments are analyzed as the psa since it is the only argument coindexed with the verb and the target of a range of syntactic operations (though non-psa actors retain several subject-like properties2, e.g., himmelmann 2008; shibatani 1991; kroeger 1993; schachter 1976; 1977; 1995). likewise, arguments that are marked by ng or sa will be referred to as the non-privileged syntactic arguments (npsa). undergoer voice structures have angmarked (psa) undergoers and actor voice structures have ang-marked (psa) actors. if a predicate has voice affixation, the semantic role of the argument that is ang-marked is overtly marked by the voice affix on the predicate (himmelmann 2008): (1) a. 'the teacher bought the book' b. 'the teacher bought a book' tagalog has more than one "transitive" construction, whereby "transitive" refers to two+ participant constructions that is used to describe one entity acting on another entity. i will focus on two constructions, the actor and undergoer voice forms which have some analogy to the active/passive alternation in languages like english, but which are functionally very different. on morphological grounds, no verbal voice form in tagalog can be considered basic, as all verbs consist of a verb stem plus a distinct voice affix. furthermore, the actor is neither demoted nor dropped as is the case with actors in passive sentences, which is taken as evidence that tagalog has a symmetrical voice system, as opposed to an asymmetrical voice system as is seen with the english active/passive patterns (e.g., latrouite 2011; schachter 1976, 1977, 1995; himmelmann 2008). 2.3. tagalog undergoer voice preference in tagalog, undergoers are the preferred privileged syntactic argument (e.g., cena, 1977; cooreman et al., 1987; wouk, 1986; garcia & kidd, 2021), which seems contrary to robust cross-linguistic patterns for actors as psa (see riesberg & primus 2015). the undergoer voice preference has been seen in narrative text (katagiri 2005; wouk 1986; cooreman et al., 1984: 17), in child speech (marzan 2013; garcia et al. 2018), and psycholinguistic experiments (tanaka et al., 2016). various semantic-pragmatic factors have been proposed to affect angmarking in tagalog, including, but not limited to topicality, specificity, definiteness, animacy, and (to some degree) verb semantics. however, apart from latrouite (2011), little work has explored the undergoer voice preference with respect to their verbs and their co-occurrence patterns with their arguments' semantic-pragmatic factors. bili ng guru ang libro buy.prv3 ng teacher ang book bili ang guru ng libro buy.prv ang teacher ng book 2.4. characteristics of ang-marked arguments basic sentences typically have one ang-phrase (schachter & otanes 1972)4. the ang-marked argument is understood as the most "prominent" (latrouite 2011) or "salient" (wouk 1986) argument. here, prominence can be understood broadly in terms of marking the argument that has the most relevance to the message or utterance (latrouite 2011; relevance theory, sperber & wilson, 2004). to ang-mark an entity is to indicate who or what the verb is about. generally, argument prominence is measured in multiple ways, including definiteness, animacy, topicality, and others5, which will be briefly described below. in tagalog, pronouns and proper names are not marked by ang, ng, or sa, however they have corresponding forms to the three markers (here, glossed as ang, ng, sa). though prominence-marking is primarily a pragmatic notion, prominence can be analyzed along some semantic scales that allows us to examine a complex weighting system between features to explain patterns in the data. definiteness and ang-marking highly definite entities6, particularly undergoers, are more likely to be ang-marked (bowen 1965; schachter & otanes 1972; naylor 1975, etc.). the role of definiteness in ang-marking seems to be exhibited in other philippine languages, e.g., ilokano (schwartz 1976), hiligaynon (wolfenden 1971) and cebuano (wolff 1966). schachter (1976) and schwartz (1976) proposed that definite undergoers will be ang-marked and in the case where none of the nominals is definite, tagalog may resort to ang-less existential constructions (schachter & otanes, 1972). however, adams & manaster-ramos (1988) show that indefinite readings of ang-marked nouns are not only possible, but the generally accepted interpretation, when there is an indefinite quantifier such as isa-ng ‘one,’ marami-ng ‘many,’ or anuman ‘anything,’ and others: (2) tawag-an ang isa-ng pediatric dermatologist kung na-pansin call-uv.imp ang one-lnk pediatric dermatologist cond uv.pfv-notice mo ang anuman-ng naaangkop ng abcde 2sg ang anything-lnk appropriate ng abcde ‘call a pediatric dermatologist if you notice something that looks like whatever is labeled in abcde (context: pictures associated with different health issues)’ (sketchengine tltenten2019, website: krikids.com) in 2, only an indefinite interpretation for ang isang pediatric dermatologist is possible despite its ang-marked status, suggesting that undergoers, regardless of definiteness can be ang-marked. furthermore, under a discourse-based definition of definiteness, definiteness appears to be weighed differently between the undergoer voice and actor voice such that highly definite undergoers tend to be ang-marked, but indefinite undergoers do not necessarily mean actors will be ang-marked (wouk 1986, see similar proposal for specific undergoers vs. actors in latrouite, 2011). animacy and ang-marking animacy, often correlated with definiteness and other referential features, also plays a role in ang-marking. generally, the actor voice (ang-marked actors) tends to be less acceptable with a human undergoer. for example, examples from saclot (2006) shows: (3) a. 'the dog bit a bone' b. 'a dog bit the bone/lena' c. ??'the dog bit me/lena' 3a,b show that actor voice and undergoer voice forms of kagat 'bite' are acceptable when the undergoer is inanimate. 3a shows that actor voice is acceptable when the undergoer is inanimate but 3b shows that it is much less acceptable when the undergoer is human. sentence kagat ang aso ng/sa buto bite ang dog ng/sa bone kagat ng aso ang buto/si lena bite ng dog ang bone/ang lena ??kagat ang aso sa akin/kay lena bite ang dog sa 1sg.sa/sa lena production experiments with tagalog speakers show that participants prefer to use the undergoer voice when the undergoer is human, even when the actor is also human and the undergoer is human or non-human (sauppe 2017). predicate-inherent orientation and ang-marking slightly less explored are these undergoer/actor voice structures with respect to verbs and their arguments. latrouite (2011, 2016) proposes an analysis that incorporates aspects of the verb's meaning along with the semantic-pragmatic features discussed above to explain asymmetrical patterns of undergoer and actor voice patterns ("voice marking gaps" latrouite 2011). for example: (4) a. intended: 'the children killed a dog.' b. 'the children killed the dog.' (cf. saclot 2005:3) 4a,b contrast with the examples in 3a-c. even though ang mga bata 'the children' has higher animacy and/or definiteness than aso 'dog', the actor voice form of patay 'kill' is generally unacceptable (except in certain constructions, e.g., focus construction). verbs takot 'frighten' and sira 'break' pattern similarly to patay. but either form is acceptable with the verb suntok 'hit' even when both arguments are referentially prominent as in 5a,b (saclot 2006: 10, cited from latrouite 2011): (5) a. 'pedro hit jose' b. 'pedro hit jose' *patay ang mga bata ng aso pfv.kill ang pl child ng dog patay ng mga bata ang aso pfv.kill ng pl child ang dog suntok si pedro kay jose pfv.hit ang pedro sa jose suntok ni pedro si jose pfv.hit ng pedro ang jose these examples suggest that in addition to definiteness and animacy, an argument can be measured based on its prominence in the predicate structure, or its centrality to the predication (latrouite 2011 p. 194). undergoer-oriented verbs such as patay 'kill,' or takot 'frighten,' sira 'destroy,' etc., highlight the state of the undergoer compared to the actor (latrouite 2011), increasing its likelihood of occurring in the undergoer voice. by contrast, activity verbs such as kain 'eat' or sulat 'write,' which allow for incremental interpretations with individuated undergoers, describe activities that profile information about the actor instead of the undergoer. this increases the likelihood that the actor will be the psa. punctual contact verbs like suntok 'hit,' kagat 'bite,' and others, which denote punctual contact between the actor and undergoer may have no clear predicate-inherent focus of attention and thus might occur readily in either voice form. the choice between actor and undergoer voice might then come down to prominence features. given the prior research, undergoer and actor voice structures are a result of a complex interplay between verbs and the referential properties of their arguments. the current corpus study is a first attempt at examining the extent to which the undergoer voice preference is informed by these factors across a range of verbs and large samples. 3. methods 3.1. data all data were extracted from the tltenten 2019 tagalog (filipino) web-based corpus which is part of the tenten corpora (jakubíček et al., 2013) and made up of web-crawled texts collected from the internet. the corpus has 198,303,250 words (jakubíček et al., 2013). the corpus was previously pos-tagged using a filipino-tagger model (go & nocon, 2017) which was previously based on the stanford parser (e.g., toutanova & manning, 2000). all collocational analyses and extractions were performed using the sketchengine concordance and corpus query language (cql) search tools which allows you to search for grammatical or lexical patterns in the corpus. the resulting concordance searches were pre-processed by the author to ensure the token was a valid sample of the target verb prior to annotation. the window for each extraction included two to three sentences preceding and following each token to provide some minimum context for the token. 3.2. verb selection latrouite (2011) provides only a few verbs in her predicate-inherent orientation categories, so most verbs examined here are not a priori categorized. verbs were selected based on their corpus frequency and their potential for denoting causatively transitive actions (kain 'eat' was included here). the top 1000 most frequent verbs were extracted, and then causatively transitive verbs were chosen for analyses. that is, selected verbs denoted actions that had causation and affectedness (e.g., hopper & thompson, 1980) and which had potential to take (at least) two participants as arguments in a clause. if they met the previous criteria and were also analyzed in latrouite (2011), they were also extracted. valid verbs under these criteria included bangga 'bump,' karga 'carry,' buhat ‘lift,’ patay ‘kill,’ kain 'eat,' and others, and excluded highly frequent experiencer-theme verbs such as kita ‘see,’ sabi ‘say,’ and others. 3.3. annotation scheme each token was annotated for a variety of features with respect to the verb and the psa and npsa. given how impoverished discourse-pragmatic information can be in this kind of corpora, important discourse-pragmatic features like topicality and specificity could not be coded for. instead, the author coded for related factors such as definiteness and animacy which can exhibit explicit morphosyntactic coding in minimal context. an abridged annotation scheme is shown below: verb voice affixes: voice affixes (if they existed since verbs can be bare) were coded for and co-indexed with the ang-marked argument. actor voice affixes included but were not limited to -um-, mag-, maka-, and others. undergoer voice affixes included -in-, -an, i-, and others macroroles actor and undergoer roles were assigned to the ang, ng, and sa arguments (if present) based on their role in the sentence. definiteness was broadly defined as a feature of a referent in which the hearer is not free to assign any value to the referent. they are often subject to a familiarity requirement where the value of a referential term is determined by previous discourse and/or context (aissen 2003) and their absence, presence, and individuation in the context (wouk 1986). definiteness codes were broadly based on a typological definiteness scale: personal pronoun > proper name > definite np > indefinite specific np > non-specific np (e.g., aissen 2003). arguments were coded as "definite" if they were a personal/demonstrative pronoun, proper name, discourse-old (bhatia et al., 2014), syntactically individuated as heads of relative clauses, nps marked by certain quantifiers, etc. (wouk, 1986). arguments were marked as "indefinite" if they were common nouns that were new in the context, preceded by quantifiers such as isang 'one,' or kahit 'any,' and others. arguments were marked as "other" if the definiteness values could not be determined or if that argument did not exist. animacy was coded following a typological animacy hierarchy: human >animate> inanimate > abstract (e.g., primus, 1999; aissen, 2003). hierarchical values for the referential properties allow us to calculate relative weights of those features between arguments and derive different measures of "prominence." furthermore, taking the verb and clausal properties into account with these weights provides us a way of understanding how these features might inform a verb's occurrence in actor and undergoer voiced structures based on the semantic properties of their arguments. 4. preliminary results a total of ten verbs and 685 tokens were annotated and analyzed here. figure 1 shows the proportions of occurrence for undergoer (green bars) and actor voice (blue bars) and "other" uses (yellow bars) for each verb in the sample. the "other" category encompassed infrequent voice forms (e.g., ang-marked locations or instruments), forms where the ang-entity was a reciprocal pronoun (e.g. kita 1sg.2sg pronoun), when the clause was ang-less, or if the clause had double-ang arguments. in general, in line with the prior research, ang-marked undergoers were more frequent than actors in the entire sample (green bars, 50.7% of all usages compared to 27.8% of all usages). figure 1. verb-specific patterns for ang-marking undergoers and actors eight of ten verbs, habol 'chase,' tawag 'call,' tulak 'push,' karga 'carry,' buhat lift,' hawak 'hold/touch,' punas 'wipe,' and patay 'kill,' occurred between 50-60% of the time in the undergoer voice. in comparison, the verbs kain 'eat' and bangga 'bump' tend to have more angmarked actors (54% and 42% respectively). the verb kain occurred in both intransitive and transitive uses, which contributes to the high proportion of ang-marked actors. the result for kain provides evidence for latrouite's (2011) analysis that kain is more actor-oriented as an incremental activity verb. the verb bangga 'bump' appears to have less of a preference for angmarking undergoers compared to the other verbs. this verb might pattern similarly to suntok 'hit,' since bangga denotes a meaning of surface contact. bangga's distribution between actor and undergoer voice structures might provide support for verbs of surface contact exhibiting "neutral" predicate-orientation. broadly speaking, there appears to be relatively little variability between verbs in how they are used in the undergoer and actor voice sentences. the following sections will further examine the extent to which the relative features of definiteness and 0.17 0.36 0.51 0.52 0.52 0.56 0.60 0.60 0.61 0.62 0.54 0.42 0.35 0.17 0.25 0.27 0.21 0.22 0.20 0.17 0.29 0.22 0.14 0.31 0.23 0.17 0.19 0.18 0.19 0.21 0.00 0.10 0.20 0.30 0.40 0.50 0.60 0.70 0.80 0.90 1.00 kain (63) 'eat' bangga (67) 'bump' habol (74) 'chase' tawag (83) 'call' tulak (75) 'push' karga (63) 'carry' buhat (47) 'lift' hawak (83) 'hold/touch' punas (84) 'wipe' patay (47) 'kill' p ro p o rt io n s o f o cc u rr e n ce verbs verb-specific patterns for ang-marking undergoers and actors undergoer actor other animacy have an influence on an individual verb's distributions in undergoer and actor voice structures. 4.1. relative definiteness relative definiteness was calculated by comparing the definiteness feature between the angargument and the ngor saargument if present. if there was no explicit ng-/sa-entity, they were coded as being absent and having a value akin to "0" in the weight calculations. to better understand how each verb might be affected by relative definiteness and its role in ang-marking actors versus undergoers, the data was further separated by the proportions of relative definiteness on undergoer voice (figure 2) and actor voice (figure 3) per verb. figure 2. relative definiteness on ang-marked undergoers per verb figure 2 shows the proportions of occurrence of undergoer voice structures where the psaundergoer had higher (blue), lower (orange), or equal (gray) definiteness to the actor. generally, undergoers and actors tended to be relatively equal in definiteness across verbs, but there is variation among the verbs in how often the undergoer was higher or lower in 0 0.1 0.2 0.3 0.4 0.5 0.6 0.7 0.8 0.9 1 hawak (50) 'hold/touch' kain (11) 'eat' buhat (28) 'lift' tulak (39) 'push' punas (51) 'wipe' habol (38) 'chase' karga (35) 'carry' patay (29) 'kill' tawag (83) 'call' bangga (24) 'bump' p ro p o rt io n s o f o cc u rr en ce verbs relative definiteness on ang-marked undergoers per verb higher lower same definiteness than the actor. verbs like patay 'kill', tawag 'call', and bangga 'bump' often appear with undergoers that are higher in definiteness compared to the actor. the event-semantics of patay 'kill' suggests that it may exhibit a natural preference towards the undergoer voice regardless of the referential status of its arguments. this is demonstrated in example 6 below: (6) 'that one (previously mentioned criminal) should also be killed' example 6 demonstrates an instance of actor absence in the undergoer voice. actor dropping can sometimes result from actor being relevant to a context such that they do not need to be referred to further. in this example, the actor is not so much "dropped" as it is just irrelevant to refer to in this context. the main focus or the most prominent entity is the entity to be killed, indicated by the ang-form demonstrative pronoun yan along with the undergoer voice marking on the verb patay-in. while bangga also occurs with undergoers higher in definiteness, the pattern for bangga 'bump' differs from what we see with patay. whereas patay exhibits an undergoer preference and argument definiteness seems to reinforce that pattern, bangga has a slight preference for the actor voice (figure 1). when it does appear in the undergoer voice, there are many examples of the undergoer with higher definiteness than the actor: (7) 'the tricycle that was being driven by the victim, clemente enerio of antipolo, was bumped/hit by a 6-wheeler truck' in example 7, the undergoer ang tricycle has higher definiteness given its further elaboration through the additional relative clause na minamaneho ng biktimang... the nature of the undergoer having higher definiteness in these bangga cases often occur due to some relativization process that provides further elaboration and emphasizes the importance of the dapat patay-in na rin yan should kill-uv.irr lnk also dem.ang na-bangga ng 6 wheeler truck ang tricycle na uv.pfv-bump/hit ng 6 wheeler truck ang tricycle rel ma~maneho ng biktima-ng si clemente enerio ng antipolo ipfvdrive ng victim-lnk ang clemente enerio gen antipolo undergoer in these instances (wouk 1986). because bangga may be a neutral verb, we might speculate that the role of relative definiteness plays a slightly more important role for undergoer uses of bangga compared to patay. however, the presence of undergoers with lower and equal definiteness to the actors suggest that this is not the entire story for bangga. that a more definite entity would be ang-marked is not surprising in and of itself. more interesting are the patterns for the verbs where undergoer and actor definiteness are equal. across verbs like hawak 'hold/touch,' buhat 'lift,' tulak 'push,' punas 'wipe,' and a couple others, undergoers and actors often were equal in definiteness (gray bar). both arguments were generally referred to using pronouns, possessed body parts, personal names, or descriptive nps in situations where there was physical contact between the arguments: (8) 'clyde held my hand' given that these verbs denote physical contact but not necessarily result-oriented action, we might analyze these verbs as having less of a predicate-inherent orientation and expect undergoers to have higher definiteness when the verb occurs in undergoer voice. instead, the general undergoer voice preference despite the equal weights potentially suggests that other factors may affect undergoer uses of these verbs or that the undergoer voice takes precedence over argument referential properties and predicate semantics. the results in figure 2 suggest that these verbs generally exhibit a preference for the undergoer voice and that relative definiteness may be a factor, but not always a defining factor of undergoers for these verbs. the co-occurrence patterns of relative definiteness in the undergoer voice contrasts heavily with the actor voice (figure 3). except for punas 'wipe,' ang-marked actors almost always have higher definiteness compared to the undergoer. this accords with prior work that shows that the tagalog actor voice is less frequent and more constrained (latrouite 2011; 2016) and perhaps more marked. the contexts for actor ang-marking may rely more on the actor having higher definiteness than the undergoer in comparison to the undergoer voice (wouk 1986). hawak-an ni clyde ang aki-ng kamay hold-uv ng clyde ang 1gen-lnk hand figure 3. relative definiteness on ang-marked actors across verbs the verb punas 'wipe' appears to be the primary exception to this. many of the undergoers in these instances were often body parts that were implicitly co-referential with the actor: (9) 'she wiped (her) mouth' both actors and undergoers in such examples exhibit equal definiteness values. but as we saw in figure 1, punas has a strong preference for undergoer voice relative to actor voice, and in figure 2 looking at the relative definiteness values, punas tends to have ang-marked undergoers even when relative definiteness was equal between the arguments. there are too few samples to draw strong conclusions, however, this may suggest that relative definiteness is not as strong a differentiating factor between ang-marking undergoers and actors for punas. 0.00 0.10 0.20 0.30 0.40 0.50 0.60 0.70 0.80 0.90 1.00 punas (17) 'wipe' karga (17) 'carry' kain (34) 'eat' buhat (10) 'lift' hawak (18) 'hold/touch' tulak (19) 'push' habol (26) 'chase' bangga (28) 'bump' patay (8) 'kill' tawag (14) 'call' p ro p o rt io n s o f o cc u rr e n ce verbs relative definiteness on ang-marked actors per verb higher lower same nag-punas siya ng bibig av-wipe 3.ang ng mouth in sum, the role of relative definiteness for ang-marking undergoers (figure 2) seems to be variable across verbs. by contrast, higher relative definiteness on actors is generally the default when the verb is used in the actor voice. certain verbs, such as punas 'wipe' may exhibit verbspecific behaviors that do not follow this. 4.2. relative animacy relative animacy was calculated by comparing the animacy features (human, animate, inanimate, abstract) between the psa (ang-marked) and the npsa (ng or sa-marked) if present. figure 4 shows the relative animacy values for verbs in undergoer voice (ang-marked undergoers). figure 4. relative animacy on ang-marked undergoers across verbs figure 4 shows that though undergoers and actors were often equal in animacy, there is variation across verbs much like relative definiteness on undergoer voice marking. the few instances of undergoer voice kain 'eat' initially seem odd. how could the undergoer (eatee) have 0 0.1 0.2 0.3 0.4 0.5 0.6 0.7 0.8 0.9 1 kain (11) 'eat' karga (35) 'carry' buhat (28) 'lift' hawak (50) 'hold/touch' habol (38) 'chase' tawag (83) 'call' patay (29) 'kill' punas (51) 'wipe' bangga (24) 'bump' tulak (39) 'push' p ro p o rt io n s o f o cc u rr en ce verbs relative animacy on ang-marked undergoers per verb higher lower same equal or higher animacy compared to the actor (eater)? upon further inspection, these were not examples of cannibalism. instead, situations of intimacy used kain less literally where the undergoer was often a body part like labi 'lips' (coded as human). other examples of when the undergoer was of higher animacy than the actor was when the actor was non-referential: (10) 'make some meat soup when needed because it is delicious to eat it when it is freshly cooked' 10 shows kain-in used with ito-ng, a demonstrative pronoun which refers to the previously mentioned sabaw ng karne 'meat soup'. although kain predicate-inherently profiles the actor, when it is absent as in 10, the undergoer must be the most prominent argument. the verb karga 'carry' seems to pattern differently from the other verbs. that is, there is a high proportion of ang-marked undergoers that are lower in animacy compared to the actor: (11) 'after i carried my things to the car, i said goodbye to dave' karga denotes an action wherein the object changes location and thus might be understood as affected. since carrying objects is a common occurrence, undergoers lower in animacy are not surprising and may potentially reinforce the undergoer voice preference. figure 5 shows the relative animacy of ang-marked actors compared to the undergoer if present. gawa ng sabaw ng karne kapag kailangan lamang sa make ng soup gen meat adv.when need adv.only sa dahil(a)-ng masarap ito-ng kain-in n(an)g bagong-luto because-lnk tasty dem.ang-lnk eat-uv.nfin adv newly-cooked matapos i-karga ang gamit ko sa kotse finished uv-carry ang things 1.gen sa car nag-paalam na ako kay dave av-goodbye adv 1sg.ang 3sg.sa dave figure 5. relative animacy on ang-marked actors across verbs like relative definiteness on ang-marked actors, there is far less variation in the role of relative animacy of actors compared to undergoers in actor voice sentences. actors tended to have higher (or at least equal) animacy values compared to undergoers across verbs, suggesting again that the actor voice appears to be more constrained such that actor voice usage tends to occur when actors are higher in prominence. this may vary with verb. the pattern for the verb karga 'carry' exhibits a complementary pattern to what we saw in the undergoer voice. that is, in the actor voice, actors frequently have equal animacy to undergoers and higher animacy does not appear to be the defining factor for ang-marked actors. this pattern along with what we see in the undergoer voice for karga (figure 4) provides further evidence that relative animacy may not have as strong a relationship with karga uses. 0 0.1 0.2 0.3 0.4 0.5 0.6 0.7 0.8 0.9 1 karga (17) 'carry' punas (17) 'wipe' habol (26) 'chase' tawag (14) 'call' bangga (28) 'bump' tulak (19) 'push' buhat (10) 'lift' kain (34) 'eat' hawak (18) 'hold/touch' patay (8) 'kill' p ro p o rt io n s o f o cc u rr e n ce verbs relative animacy on ang-marked actors per verb higher lower same the actor voice uses of patay 'kill' bear examining. as previously mentioned, under an eventstructural analysis, patay highlights the resultant state of the undergoer which contributes to its preference for the undergoer voice. the actor voice examples found here show that when patay occurs in the actor voice, those actors have high definiteness (figure 4) and animacy (figure 5). however, it may be more accurate to characterize these occurrences as when the undergoer is non-referential (and therefore not definite/animate): (12) '…when you've been able to kill in order to protect yourself or your family, that is just' 12 shows the actor ka 'you' as the only argument to the verb. there is no undergoer explicitly or implicitly mentioned in this example. this is further co-indexed by the makaactor voice affix on patay. we see another example of undergoer absence in 13: (13) go signal cla (sila) go signal 3pl.ang 'yes, they do not kill, but they have a go signal…' again, the only argument to patay is the pronoun cla7 (sila) 'they,' which is co-indexed by the actor voice affix -um-. although patay's semantics licenses and perhaps profiles the undergoer, the absence of an undergoer in these examples suggests a pattern similar to english existential null complementation constructions where the patient argument of an accomplishment verb is omissible (goldberg 2005; michaelis 2006). what this shows is that patay exhibits highly specific behaviors in the actor voice that involve referential features as well as larger constructional patterns. in sum, these results show that relative animacy may play a greater role in the actor voice compared to the undergoer voice across verbs. furthermore, each verb's event …kapag naka-patay ka para proteksyunan adv abil.av.pfv-kill 2sg.ang prep protection ang sarili mo o ang pamilya mo, that is just ang self 2sg.gen conj ang family 2.gen that is just oo hindi cla (sila) pa~patay pero may yes neg 3pl.ang ipfv~kill but exist semantics interact with the relative animacy between its arguments, though there is greater variation in its co-occurrence with undergoer voice compared to the actor voice. 5. discussion & conclusions this corpus study provided a preliminary understanding of how verbs and the referential features of their co-occurring arguments interact to produce verb-specific patterns with respect to the undergoer voice and the actor voice. the data showed that across most verbs, there was a preference for the undergoer voice. an examination of the undergoer voice and actor voice show that relative definiteness and animacy of the verbs' co-occurring arguments is more indicative of the patterns for actor voice compared to the undergoer voice. that is, ang-marked actors occur more frequently when actors are more prominent in terms of definiteness and animacy compared to the undergoer. otherwise, the undergoer voice appears to be the default for most verbs, which is in line with prior work on tagalog (latrouite 2011; 2016; wouk 1986; katagiri 2005; cooreman, fox, & givón, 1984: 17). however, an examination of each verb's patterns of occurrence in both voices and the semantic properties of their arguments showed that each verb exhibited its own behaviors within these broader general patterns. in sum, there is a complex interplay between verbs, their semantics, the referential properties of their arguments, and other factors which shed light on these distributions of undergoer and actor voice structures. this study is just a start, but many more verbs and samples must be annotated and analyzed to better understand these patterns. furthermore, ang-marking is likely influenced by other factors, such as different construction types (focus constructions, e.g., latrouite 2011; null complementation; other constructions, garcia & kidd 2021) and of course, discourse factors. an important note here is that these results reflect co-occurrence patterns and not patterns of causation. that is, the presence of one factor with a certain voice marking shows that they happen to co-occur together. we cannot establish causal or directional links between these features and voice marking. if we are to better understand a causal link between these factors and what "triggers" how these forms are used, other methods such as experiments, must complement the corpus methods. references adams, k. l., & manaster-ramer, a. (1988). some questions of topic/focus choice in tagalog. oceanic linguistics, 27(1/2), 79. https://doi.org/10.2307/3623150 bowen, donald j. (ed.) 1965. beginning tagalog. berkeley/los angeles: university of california press. bresnan, j., cueni, a., nikitina, t., & baayen, r. h. (n.d.). predicting the dative alternation. 33. bybee, j. l. (2006). from usage to grammar: the mind’s response to repetition. language, 82(4), 711–733. https://doi.org/10.1353/lan.2006.0186 cena, r. m. (1977). patient primacy in tagalog. lsa annual meeting, chicago, 28–30. colleman, t. (2009). verb disposition in argument structure alternations: a corpus study of the dative alternation in dutch. language sciences, 31(5), 593–611. https://doi.org/10.1016/j.langsci.2008.01.001 cooreman, a., fox, b. a., & givón, t. (1984). the discourse definition of ergativity. studies in language, 8(1), 1–34. https://doi.org/10.1075/sl.8.1.02coo foley, w., & van valin jr, r. (1977). on the organization of “subject” properties in universal grammar. annual meeting of the berkeley linguistics society, 3, 293. https://doi.org/10.3765/bls.v3i0.3297 garcia, r., roeser, j., & höhle, b. (2018). thematic role assignment in the l1 acquisition of tagalog: use of word order and morphosyntactic markers. language acquisition, 1–27. https://doi.org/10.1080/10489223.2018.1525613 goldberg, adele. 1995. constructions: a construction grammar approach to argument structure. chicago: university of chicago press. gries, s. th., & stefanowitsch, a. (2004). extending collostructional analysis: a corpus-based perspective on `alternations’. international journal of corpus linguistics, 9(1), 97–129. https://doi.org/10.1075/ijcl.9.1.06gri himmelmann, n. p. (2008). lexical categories and voice in tagalog. voice and grammatical relations in austronesian languages, 247–293. hopper, p. j., & thompson, s. a. (1980). transitivity in grammar and discourse. language, 56(2), 251–299. https://doi.org/10.1353/lan.1980.0017 https://doi.org/10.2307/3623150 https://doi.org/10.1353/lan.2006.0186 https://doi.org/10.1016/j.langsci.2008.01.001 https://doi.org/10.1075/sl.8.1.02coo https://doi.org/10.3765/bls.v3i0.3297 https://doi.org/10.1080/10489223.2018.1525613 https://doi.org/10.1075/ijcl.9.1.06gri https://doi.org/10.1353/lan.1980.0017 katagiri, m. (2005). voice, ergativity, transitivity in tagalog and other philippine languages: a typological perspective. the many faces of austronesian voice systems. pacific linguistics. kroeger, p. r. (1993). another look at subjecthood in tagalog. 15. latrouite, a. (2011). voice and case in tagalog: the coding of prominence and orientation [dissertation]. https://d-nb.info/106308511x/34 latrouite, a. (2016). shifting perspectives: case marking restrictions and the syntaxsemantics-pragmatics interface. in j. fleischhauer, a. latrouite, & r. osswald (eds.), explorations of the syntax-semantics interface (pp. 289–318). de gruyter. https://doi.org/10.1515/9783110720297-011 marzan, jocelyn c. 2013. spoken language patterns of selected filipino toddlers and preschool children. diliman, quezon city: university of the philippines diliman dissertation. michaelis, l. (2006). complementation by construction. annual meeting of the berkeley linguistics society, 32(1), 529. https://doi.org/10.3765/bls.v32i1.3461 nagaya, n. (2006, january). topicality and reference-tracking in tagalog. in 9th philippine linguistics congress. quezon city: university of the philippines diliman. naylor, p. b. (1975). topic, focus, and emphasis in the tagalog verbal clause. oceanic linguistics, 14(1), 12. https://doi.org/10.2307/3622792 reid, lawrence and paul schachter. "tagalog." in the world's major languages (2nd edition), edited by bernard comrie, chapter 47. london: routledge, 2008. riesberg, s., & primus, b. (2015). agent prominence in symmetrical voice languages. stuf language typology and universals, 68(4). https://doi.org/10.1515/stuf-2015-0023 saclot, m. j. (2006). on the transitivity of the actor focus and patient focus constructions in tagalog. tenth international conference on austronesian linguistics, palawan, philippines, january, 17–20. sauppe, s. (2017). word order and voice influence the timing of verb planning in german sentence production. frontiers in psychology, 8. https://doi.org/10.3389/fpsyg.2017.01648 sauppe, s., norcliffe, e., & konopka, a. e. (n.d.). dependencies first: eye tracking evidence from sentence production in tagalog. 6. https://d-nb.info/106308511x/34 https://doi.org/10.1515/9783110720297-011 https://doi.org/10.3765/bls.v32i1.3461 https://doi.org/10.2307/3622792 https://doi.org/10.1515/stuf-2015-0023 https://doi.org/10.3389/fpsyg.2017.01648 schachter, paul. 1976. ‘the subject in philippine languages: topic, actor, actor-topic, or none of the above.’ in li, charles (ed.) subject and topic. new york: academic press, 493518. schachter, p. (1977). reference-related and role-related properties of subjects. grammatical relations, 279–306. https://doi.org/10.1163/9789004368866_012 schachter, paul. 1996. ‘the subject in tagalog: still none of the above.’ ucla occasional papers in linguistics no.15. los angeles: university of california. schachter, paul & fe otanes. 1972. tagalog reference grammar. berkeley: university of california press. shibatani, masayoshi. 1991. ‘grammaticization of topic into subject.’ in traugott, elizabeth c. & bernd heine (eds.) approaches to grammaticalization, vol.1. amsterdam/philadelphia: john benjamins, 93-133. tanaka, n. (2016). an asymmetry in the acquisition of tagalog relative clauses. 177. valin, r. d. v. (2009). a brief overview of role and reference grammar. 30. valin, r. d. v. (2005). a summary of role and reference grammar. 30. van valin, r. d. & lapolla, r. j. (1997). syntax: structure, meaning, and function. cambridge university press. wouk, f. (1986). transitivity in batak and tagalog. studies in language, 10(2), 391–424. https://doi.org/10.1075/sl.10.2.06wou wolfenden, e. p. (2019). hiligaynon reference grammar. university of hawaii press. wolff, j. u. (1966). beginning cebuano, part 1. yale linguistic series, 9. https://doi.org/10.1163/9789004368866_012 https://doi.org/10.1075/sl.10.2.06wou endnotes 1 there is some controversy in how to understand and gloss these markers. they've been variously understood as case markers (latrouite 2011), ang as a "topic marker" (cooreman et al., 1984), "trigger" (wouk 1986), as well as subject/object, etc. because the analysis of what the markers are is not a main issue in this paper, and there is not consensus on how to understand these markers, i will attempt to stay close to the language phenomena and just refer to them as ang, ng, and sa marking. 2 in tagalog, different subject-like behavioral properties (keenan, 1976) are distributed between the actor and the psa (ang-marked argument). for example, the npsa (non-ang-marked) actor retains many subject-like properties, such as reflexive binding, control of an actor gap in the second coordinated clause, deletion in imperatives, deletion in the second coordinated clause, and control of a gap in subordinated clauses (schachter, 1977; shibatani, 1991; kroger, 1993; shibatani, 2005; latrouite, 2011). on the other hand, undergoer psa arguments show several subject properties such as verb agreement, extractability, control of floating quantifiers and gaps in samptan ‘while’ clauses (shibatani, 1991). these behaviors have resulted in tagalog being classified as varying systems, including, but not limited to, a "focus” system (e.g., schachter & otanes, 1972; schachter, 1976; naylor, 1995, etc.) or a “trigger” system (e.g., schachter, 1976; fox, 1982; wouk, 1986). 3 abbreviations: 1 first person, 2 second person, 3 third person, abil abilitative, adv adverb, ang ang marker/form, av actor voice, conj conjunction, exist existential, gen genitive, ipfv imperfective, irr irrealis, lnk linker, neg negation, nfin non-finite, ng ng marker/form, pfv perfective, pl plural, prep preposition, rel relativizer, real realis, sa sa marker/form, sg singular, uv undergoer voice 4 there are some structures that have double-ang or only ng-marking, which are beyond the scope of this paper. 5 the nature of the corpus data does not allow for topicality to be reliably measured here, but i would be remiss in not briefly mentioning the literature around ang-marking and topicality. the concept of "topic" and "focus" have referred to different functions in tagalog, resulting in varying analyses of the ang-argument, only a couple of which i will cover here. nagaya (2006) defined a topical referent as a complex feature that denotes an "animate participant and/or an s or a core argument which tends to be referred to by a pronoun" and a non-topical referent as an "inanimate participant and/or an o core argument" which tends to be marked by zero anaphora (nagaya, 2006, p. 6). cooreman et al., (1984) measured topicality by looking at referential distance, topical persistence, and deletability. the researchers found that in "transitive" -inclauses (i.e., undergoer voice) patient arguments had much lower topicality compared to agents in terms of anaphoric and cataphoric behaviors. that is, the ng-marked actor argument was shown to be more topical than the ang-marked patient argument. in sum, the role of topicality on angmarking is complex and depending on how topicality is defined, may result in different analyses of ang-marked and ng-marked arguments. 6 although the notions of specificity and definiteness are formally separate features, in the tagalog linguistics literature, the two features have been used nearly interchangeably. 7 the form cla to mean the pronoun sila is common in internet usage for this pronoun. microsoft word kosse-cril2021-proof_final.docx 1 pull a [proper name] maureen kosse university of colorado boulder this paper considers the pull a proper name (papn) construction in english. the bulk of onomastic research in linguistics present proper names as a word class with ‘unique reference’, often comparing them to deictic expressions (cf. searle 1969). unlike deictics, however, names are interpretable beyond the immediate linguistic context. accounts from sociocultural and cognitive linguistics dispute the notion of unique reference, instead arguing that proper names vary in everyday use. proper names typically invoke specific persons; however, the data provided here indicates that names are frequently used as metonymic framing devices for specific events, generic scenarios, and hypothetical figures of personhood (agha 2007, dancygier 2011, ainiala and östman 2017). using examples from twitter, this preliminary analysis compares tokens of pull a britney [spears] and pull a karen along their constructional and conceptual qualities. while tokens of pull a britney evoke a specific person and event in time (spears’ well-known mental breakdown in 2007), tokens of pull a karen are generic in nature and index a broad array of attitudes, personality traits and behaviors. the findings of this paper support dancygier’s (2011) claim that onomastic study should center the constructional qualities of proper names as used in real-life examples from discourse. keywords: naming, onomastics, constructional compositionality, proper names 1. introduction the data for this project was collected from the corpus of global web-based english (glowbe) and supplemented by tokens from the twitter search api. the pull a [proper name] (papn) construction generally means “to behave in a similar manner to pn.” the construction is used to draw similarities between the speaking context and the qualities or events associated with the proper name. the pattern is idiomatic for two reasons: (1) its meaning cannot be identified by the composite meanings of its parts and (2) the required proper noun adds to the idiomaticity of this construction. in the examples below, understanding the utterance requires prerequisite knowledge of the entity named by a proper noun. some examples have additional context cues indicating what sort of behavior is associated with a given name, while others do not. in context, (1) belongs to a discussion about the behavior of late republican politician john mccain. during the 2008 us presidential election against barack obama, critics lambasted mccain for failing to reach out effectively to his base. without understanding john mccain’s political history, it would not be possible to understand don’t pull a john mccain as ‘do not become complacent.’ colorado research in linguistics, volume 25 (2021) 2 table 1 1 don’t pull a john mccain. 2 pull a john galt, ditch, and this will speed up the decline. 3 did he pull a john wayne, or a dirty harry when he came out with guns ablaze? 4 he has made a commitment to me that he will not pull a michael ross. 5 are you trying to pull a michael mann? 6 you need to pull a michael jordan and come back for a third time. 7 [americans] are not leaving unless the iraqis basically pull a iran on washington i found papn in the oxford english dictionary online, under the entry pull, v. below, i include the senses of pull that i believe to be most relevant. the action of pulling refers to drawing something closer to oneself by one’s own strength (in earliest known usage pull was used for feathers, hair, fruits and vegetables). table 2 pull, v. 7a. to say or do (something) with intent to deceive, or to impress or shock, etc.; also with on. 1894 ‘pull the sick list’…to get on the sick list when not ill. 1915 don’t pull any of that dope on me. 1937 not that i think anyone would pull the same trick twice. 7b. to behave in a manner characteristic of or associated with (the person specified). 1911 strunk pulled a ty cobb on henry in the seventh, scoring from second. 1931 to ‘pull a lindbergh’ means to do something heroic, but to ‘go lindbergh’ means to get the flying fever in a rather callow manner. 2004 worried that he’d pull a hendrix and choke on his own vomit, fitch rolled him onto his side. 7c. to make (a foolish mistake), to perpetrate (a blunder). 1913 [he] got his signals mixed and pulled a boner. the definition provided by 7b. describes most instances of papn, but i believe that all three of the senses i present here are relevant to our understanding of this construction as a negative stancetaking device (du bois 2007; jaffe 2009). while these definitions are helpful in assessing pull a [proper name] 3 the overall meaning of papn, we must also examine the construction in use. here are a few more examples from my data: table 3 8 bento should have probably pulled a spain and played without a striker. 9 pull a kim k n force them ""##$$%%&&''(( 10 pull a vp and fade into obscurity 11 they need to pull a wwe and just make this shit $10 a month 12 i thought buttigieg was gonna pull a beto. why is he still here? 13 he is backpedaling so hard he is going to pull a superman 2 and turn the planet backwards in function, papn is used to indirectly compare two scenarios: that of the subject referent and the qualities/events/personality of the proper name referent. (12) compares the pete buttigieg presidential campaign to that of beto o’rourke. in this example, pull a beto means ‘to withdraw from candidacy.’ in many cases, papn is followed by an elaborating clause introduced by and. 2. theoretical framing for the purpose of brevity, i frame this paper using three sources which best represent my analytical approach to papn. there is already a large existing body of onomastic work that crosses into pragmatics and language philosophy that links names to deixis and to definiteness (cf. searle 1969; kripke 1980). in the grammar of names (2007) anderson challenges this conceptualization: unlike deictics, names are not dependent on the immediate non-linguistic context. but, of course, again unlike deictics, the use of a name like basil for identification presupposes that the speaker and addressee have participated, together or separately, in a naming to them, as basil, of the same entity, and that, if separate namings are involved, they have ascertained that their namings correspond (217). in other words, the acts of identifying and naming a referent are socially informed and collaborative. using similar argumentation, ainiala and östman (2017) advocate for what they call the “socio-onomastic” approach to naming. they note that “traditional” onomastics most often analyze from either a diachronic or a typological perspective. in both of these analyses, scholars colorado research in linguistics, volume 25 (2021) 4 attempt to track the structure and trajectory of a name across space, time, and language groups. ainiala and östman focus instead on the everyday use and variation of names in discourse, with the intent to bring onomastics into conversation with contemporary sociolinguistics. the authors point out that, like other words and word-like objects, proper names are variable and have different associations according to place, time, and community (ainiala and östman 2017). most crucially, i rely on dancygier (2011). dancygier argues that contrary to prior analysis interprets proper names (pns) as having “unique reference,” meaning that proper names have traditionally been defined as “specialized pointers to objects, locations, or people in the actual world” (2011:208). according to dancygier, traditional grammars note that proper nouns do not typically take articles or modifiers which is taken to indicate their ‘unique reference’. noting that traditional grammars tend to lack real-life examples from discourse, dancygier challenges the distinctness of proper names from common nouns by examining their constructional properties. inspired by turner and fauconnier’s blending theory of constructions, dancygier argues for a construction-specific mechanism called constructional compositionality (dancygier and sweetster 2005). as work in blending theory has suggested, constructional forms may appear in nonprototypical contexts, and as such contribute to the meaning of the utterance in ways relying on selective projection, rather than only appearing in fully-profiled constructions (dancygier 2011:209). 3. comparative analysis: britney and karen in this section, i show that pull a [proper name] has two distinct construals: one more specific, and one more generic. anderson (2007) argues that proper names can land on a cline of genericness for utterances that ranges in specificity (228). anderson notes that the cline is quite fuzzy1 and that an utterance may be understood as more specific/generic in relation to speaking context. here, i provide examples scaling from most specific (a) to most generic (d): table 4 a merkel is a stinky, misbehaved cat. definite, most specific b [pointing at a cat] that cat is chunky. definite, specific c a long time ago, there were saber-toothed cats indefinite, non-specific but tensed (were) which adds specificity d a cat is an intelligent animal. indefinite, most generic pull a [proper name] 5 while analyzing data for papn, i noticed a great deal of variation when it came to referent specificity. i made note of pns that appear most frequently in my data: britney [spears], [donald] trump, [bill] clinton, and karen. to preserve my own sanity, i chose britney and karen as my case studies. to pull a britney [spears] means something akin to ‘have a meltdown’ or, more specifically, ‘have a meltdown and shave one’s head’. spears’ surname rarely appears in the data; i believe it speaks to the iconicity of britney spears as a performer and the massive impact her 2007 public breakdown/liberatory head shaving had on pop culture. in a sense, the indefinite article is misleading; after all, we know this isn’t ‘a’ britney, it’s ‘the’ britney. this leads me to think that pull a britney isn’t about britney spears alone; instead, i will pull a dancygier (2011) and argue that, in these examples, the pn britney metonymically represents the entire 2007 britney spears mental health crisis event. figure 1. the one and only britney spears table 5 pull a britney 14 please don’t pull a britney though! 15 she didn’t pull a britney and shave her head in front of the paparazzi. 16 don’t make me pull a britney 2007 17 how does it feel knowing somewhere out there one of your fans would probably pull a britney spears 2007 just to make you laugh? colorado research in linguistics, volume 25 (2021) 6 in these examples, britney has a specific definite referent (britney spears). britney spears hit superstar status in the pop music industry in the 90s, and by 2007 spears’ public persona suffered under the misogynistic panopticon of us popular culture, faced with constant paparazzi harassment as well as abuse from family members, partners, and the music industry at large. while the causes of this event are too manifold to outline for the purpose of this analysis, this ultimately culminated in an infamous public breakdown during february 2007, wherein spears shaved her head. in all of the examples (and in all of the instances of britney in my dataset), pull a britney is used to refer to a public meltdown. it interests me that the oed sense of papn does not seem to adequately cover the use of pull a britney. yes, these examples could be understood to mean “to behave in a manner similar to britney spears,” yet all of the instances refer to one specific event. this reading is further emphasized by (16) and (17), which both use 2007 to further specify britney. all instances of pull a britney actually refer to this one event, metonymically represented by the name britney. as mentioned in the introduction, speakers use papn to compare scenarios between the subject referent and the pn referent. table 6 18 i’m about to pull a britney and shave my head. i’m on that level today. 19 waiting for a final grade to be put into blackboard while you have an 88.5 in a class? same. i’m about to pull a britney spears i’m so stressed 20 don lemon is having such a meltdown i’m just waiting for him to pull a britney spears. in (18) and (19), both speakers compare their own emotional state to that of britney spears; in (20), the speaker overtly links don lemon’s alleged meltdown to this same event. i argue that the definite and specific qualities of the pn britney and the notoriety of her mental health struggles facilitate a construal that focuses on events. in contrast, i believe that more generic pns facilitate a construal focused on the stereotypical personality traits and behaviors, as in pull a karen. in online use, the pn karen does not have a particular referent; rather, karen refers to a stereotypical figure of personhood, “socially recognizable personae that can be performed through semiotic enactment” (agha 2007). karen is not the only generic pn used this way in the data; pns like becky appear frequently as well (though becky denotes a different kind of persona). the meme below, called the “karen starter pack,” we can see some of the semiotic resources associated with the karen persona: pull a [proper name] 7 figure 2. “the karen starter pack” (https://knowyourmeme.com/photos/1506963-karen) twitter data of pull a karen shows how speakers imagine the behavior and attitude of a “karen”: karens are demanding, condescending, entitled, and confrontational. most importantly, karen is the type of person to go over someone’s head and resolve conflict by invoking some higher authority e.g. store managers or the police. table 7 pull a karen 21 i’m about to be that bitch who complains that they didn’t get priority boarding. let me pull a karen rn. 22 don’t pull a karen and call me dear sweetie 23 pull a karen and call the police wtf 24 gonna pull a karen real quick, but legit fuck companies that won’t refund you no matter what. 25 tomorrow i have to go pull a karen and demand my whole new set of nails to be redone because they’re trash i’m anxious. 26 pull a karen and get him fired 27 finna pull a karen and speak to the manager2 colorado research in linguistics, volume 25 (2021) 8 while pull a karen functions similarly to pull a britney, the karen examples do not refer to an event but rather how a “karen”-type person might approach the scenario at hand. there is no specific referent called karen; in these examples, karen is an abstract social figure based on stereotypes of middle-aged wasps/soccer moms/suburbanites. this brings me to the important part of the analysis: how important is genericness to papn? papn functions similarly across both the britney and the karen sets, but i argue that papn can access different scales of comparison relative to the specificity of the pn. in the case of pull a britney, the name britney profiles the larger frame event (the 2007 britney mental health crisis). dancygier (2011) argues that proper names should be understood in terms of their “specific and rich” framing (209). contrary to analyses of proper names which merely writes them off as ‘nouns with unique reference’, dancygier states that rich, complex frames guide discourse such that only one referent can fit the frame evoked. drawing from the idea of constructions as blends (fauconnier and turner 2002), dancygier and sweetser (2005) propose a mechanism called constructional compositionality: the concept reflects various observations suggesting that construction specific forms (such verb forms) may appear in contexts other than fully-profiled constructions and contribute to the overall meaning in ways relying on selective projection (as described in blending theory), rather than on the additive mechanisms of compositional semantics (dancygier 2011:209). constructional compositionality relies on frame metonymy. frame metonymy describes a usage in which one aspect of a frame is used to evoke the entire frame. a proper name may evoke certain frames which influence its construal in discourse (dancygier 2011:212). for example, when a diner at table 3 is ready to pay, a server may say to another, “table 3 wants her check.” customers and table numbers are closely related within the ‘restaurant’ frame that a table number can metonymically represent the customer. an utterance like table 3 wants her check makes sense between two servers in a restaurant but not between two mathematicians for whom table 3 might evoke a different frame entirely. pull a britney is a quintessential representation of frame metonymy. every single instance of pull a britney was, in fact, a reference to an event that built up over the course of years, finally culminating in a drastic public spectacle that defines her career even a decade later. britney herself pull a [proper name] 9 is not only the center of the event, but an event so iconic that her name evokes a hyperspecified frame. on the other hand, we have the karen data, where karen is a hypothetical figure of personhood linked to a generic scenario (unnecessarily invoking conflicts and/or authority figures). it is remarkable to me that pull a karen is used so similarly by unrelated speakers. in my opinion, this suggests that karen might be undergoing some level of grammaticalization. both of these examples show the exact same construction (papn). in theory, dancygier’s analysis should account for both pull a britney and pull a karen, but i am not convinced that it does. pull a britney evokes an obvious, singular reading because the 2007 britney spears meltdown is deeply woven into american pop culture. but there is no one karen, and instead the name karen metonymically evokes a frame concerning the behaviors and traits of middle-class white women (ie. a figure of personhood). the contrasts explored in this section reflect what anderson (2007) calls the cycle of individualization (236): figure 3. cycle of individualization (anderson 2007) in future work, i would like to revisit how genericness and individualization of proper names affect construal in the papn construction. 4. idiomatic properties in this section, i analyze papn constructionally, drawing from the paradigm set by croft and cruse (2004). while normally i am far more interested in verbs, i believe that the pn in this construction is the source of its rich idiomaticity. colorado research in linguistics, volume 25 (2021) 10 table 8 conventionality interpretation of papn requires additional, complex knowledge concerning the pn referent. inflexibility papn can take various tense and aspect markers, but the pn referent does not seem to take modification other than an indefinite article. *a britney was pulled by my sister ?tomorrow i’ll pull a big karen figuration papn is a metonymic device in which proper names profile larger frames of reference (dancygier 2011). the pn serves as an emblem or metonymic standin for the event or traits associated with the pn referent. proverbiality anderson (2007:222) writers, “rather obviously, use of a name for identification presupposes prior nomination.” papn requires the hearer to have some knowledge concerning the pn referent. informality informal3 affect papn is most frequently used to draw a negative comparison between two similar scenarios. the data also indicates that papn stacks with the malefactive on construction, e.g. don’t pull a trump on the american people. knowledge of the pn referent is also an example of indexical competency or ‘know how’ (silverstein 2003). 5. conclusion: future directions while gathering twitter data for this project, i noticed three other pull constructions that, while not strictly papn, have a similar structure and meaning. i am especially interested in the third construction, pull a “quote.” in examples 31-34, the quoted text metonymically evokes the context in which one might habitually hear such an utterance. i wonder if the quotation in this construction is analogous to the proper name in papn? it is a highly idiomatic way to reference a situation, and is worth further examination. pull a [proper name] 11 table 9 x needs to pull a page out of y’s book 28 kirby needs to pull a page from dabos book. those boys get away with everything. including but not limited to [various malefactive activities] pull a [political] card 29 yeah plz make sure they don’t pull a liberal card on this one 30 you soft ass bammas are all the same. get in your feelings and you pull a race card... )*+,-. pull a “quote” 31 lmao in-laws are good pretenders shame. they will like you until you get comfortable, just when you're comfortable they will pull a "don't think you know somebody" //001122 lmao. 32 when shells n david pull a „we got scammed and don’t have tix“ to get in somehow 33 if he literally drops [a song] on christmas day, i’m gonna pull a “fuck it” and cover it for his birthday 34 i live within 30 miles of a nuclear reactor as well so don't try to pull a "but my community" argument. there other elements that i treat as uncontroversial in this paper, but interest me on a broader level. for instance, i return to example (13) from earlier: 13 he is backpedaling so hard he is going to pull a superman 2 and turn the planet backwards the papn construction is far more complex than i initially assumed. papn relies on metonymy as its primary mechanism. speakers need to have a great degree of shared knowledge in order to correctly identify a proper noun referent (and thus the evoked frame). there were many instances of the papn construction that made absolutely no sense to me because i do not listen to korean pop music or watch televised sports. are there any meaningful differences between human proper names and other types of proper nouns like titles? what about corporate entities, as in will disney pull a netflix? i think that this avenue of inquiry has a lot of potential, especially to see how well metaphors gain their specific construal via different constructions. papn is also used to dramatic social effect. it demonstrates a negative evaluative stance on the part of the speaker as they draw an unfavorable comparison between the syntactic subject and the pn referent in the predicate. this study also raises questions on reference, iconicity, and proper names from a diachronic perspective; how, over time, do proper names come to be associated with specific social colorado research in linguistics, volume 25 (2021) 12 dimensions? this would be an excellent opportunity for cognitive and sociocultural linguistics to come into this conversation, as it is increasingly clear how each enriches the other. references agha, a. (2007). language and social relations. cambridge: cambridge university press. ainiala, t., & östman, j-o. (2017).socio-onomastics: the pragmatics of names. amsterdam: john benjamins publishing company. anderson, j. (2007). the grammar of names. oxford: oup oxford. barcelona, antonio (2011). properties and prototype structure of metonymy. in r. benczes, a. barcelona, & f. j. ruiz de mendoza ibanez (eds.), defining metonymy in cognitive linguistics (7–50). amsterdam/philadelphia: john benjamins. croft, w. (2006). on explaining metonymy: comment on peirsmanand geeraerts. cognitive linguistics 17(3), 317-326. croft, w. & cruse, a. (2004). cognitive linguistics. cambridge: cambridge university press. dancygier, b. (2011). modification and constructional blends in the use of proper names. constructions and frames 3(2), 208-235. dancygier, b. & sweetser, e. (2005). mental spaces in grammar: conditional constructions. cambridge: cambridge university press. du bois, j. (2007). the stance triangle. in r. englebretson (ed.), stancetaking in discourse. philadelphia: john benjamins. fauconnier, g. & turner, m. (2002). the way we think: conceptual blending and the mind’s hidden complexities. new york: basic books. jaffe, a. (2009). stance: sociolinguistic perspectives. oxford university press. kripke, s. (1980). naming and necessity. cambridge: harvard university press. searle, j. (1969). speech acts: an essay in the philosophy of language. cambridge: cambridge university press. silverstein, m. (2003). indexical order and the dialectics of sociolinguistic life. language & communication 23, 193-229. sullivan, k. (2013). frames and constructions in metaphoric language. philadelphia: john benjamins. pull a [proper name] 13 endnotes 1 while anderson frames this as a potential problem with the model, i think its indistinct boundaries better reflect the variation we observe in casual speech. 2 note the modal finna, used in african american english; karen crosses dialects! 3 small caveat: since twitter is arguably an informal speech area, this is an informed guess. “ownself check ownself”: the role of singlish humor in the rise of the opposition politician in singapore “ownself check ownself”: the role of singlish humor in the rise of the opposition politician in singapore velda khoo university of colorado boulder the people’s action party (pap) have won every election in singapore since 1959 when the citystate was first granted self-governance. over the years, its regime has been described as authoritarian by political observers (rodan 2004; tan 2012), the subjugation of the media a commonly brought-up example of the party’s ability to shut down contrasting political views (seow,1998). with media laws that dictate the freedom of the press and protect the pap’s interests, opposition parties have found it difficult to break their stronghold on the nation-state, and there has been no real political contestation in the general elections. since 2011 however, the pap, amidst social pressure to ‘keep up with the times’, have cautiously lifted the total ban on online campaigning and as a result, singapore politics have undergone rapid mediatization. this has led to two major changes in the local political arena. firstly, the shift in symbiotic relationships between the mainstream media, political organizations and the electorate in singapore, has encouraged the paralleled rise of "newly competitive" opposition parties able to capitalize on newer, non-traditional spaces of communication to question the ruling legitimacy of the pap (ortmann, 2010). in order to brand themselves as alternative voices to an elite pap, their public performances have appealed to growing populism, and tap on singlish, an ideologically valuable linguistic resource, to do so. this paper analyzes the creative, patterned use of singlish, indexically tied to "the common singaporean" (j. leimgruber, 2013), by opposition politicians in rallies to humorously attack pap candidates and ideas. i argue that such linked uses of humor to language allow for opposition politicians to simultaneously position themselves as fellow lay members of the singaporean community, and reinforce their own political stances through the deriding of the ruling party. secondly, the rise of social media and alternative new media on the internet have created an increasingly sophisticated citizenry (cf. mazzoleni and schulz, 1999) that exhibit greater degrees of "open political dissent" (ortmann, 2010) and scrutinize political actors closer than before. this paper tracks online singlish memes in which singaporean netizens mock the pap’s ’inauthentic’ expressions of singaporean-ness and legitimize opposition politicians’ use of the language. as such, an alternative linguistic marketplace (bourdieu, 1977) emerges in which singlish humor is a symbol of populist resistance and solidarity. through the analysis of these metalinguistic commentaries, i make a case for the commodification of singlish as an ideological resource through which singaporeans construct intersubjectivity and discuss how the nation-state is aligned with certain ways of using language. keywords: sociocultural linguistics, singlish, language and identity 1 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 1. introduction on 2 september 2015, nine days before the 2015 singapore general elections, the workers’ party (wp), an opposition political party in singapore, held its first, highly-anticipated election rally of the season. five years before in the previous general elections, the wp garnered six parliamentary seats out of the total 87 seats contested, a shock breakthrough in historically “oneparty singapore”, a country which has been ruled since independence by the people’s action party (pap). six seats out of 87 was just under 7% of all seats up for election: yet this was seen as a groundbreaking feat, considering that this gave the opposition its largest representation in parliament in the history of the state (tan 2014). on that evening of 2 september, more than 50,000 people (cochrane 2015) gathered in an open field to listen to the wp, a number far outnumbering the 1000+ crowd at the pap rally on the same night. aerial photographs of the massive congregation spread quickly online, the most viral being user-generated memes where scenes of the wp crowd were juxtaposed next to pictures of the considerably smaller pap rally crowd (see figure 1). figure 1. graphic image from sgag post on 2 september 2017 (sgag 2015) 2 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 a few speeches into the rally, popular wp member-of parliament pritam singh gets to the podium and begins his rally speech. he references pap emeritus senior minister goh chok tong’s statement that the pap did not need opposition parties in parliament to act as checks and balances “we are our own checks”, goh had declared in an interview just a week ago (philomin 2015). singh denounces this together with a rhetorical question in his speech. (1) singh: mr goh chok tong himself, famously said, that the pap exists in a political system where i quote, they are their own check! crowd: ((rally crowd boos)) singh: is this the future we want for singapore or our children in the next fifty years? crowd: no! singh: ownself check ownself! crowd: ((laughter and applause)) singh derides the pap’s claim, implying that performing their own checks and balances is an illogical expectation for the electorate. at this point in the speech, he deviates from what is recognizable as standard singapore english, and uses an innovative singlish expression1, the contact vernacular spoken by many singaporeans. the closest english translation of “ownself check ownself” would be something like “on your own you check yourself”. this brief, humorous digression in an otherwise completely serious speech catches the attention of the public. overnight, singh’s “ownself check ownself” circulates widely across many political websites and social media platforms. the 31-second video segment transcribed above is posted on youtube with the title “pritam singh: pap ownself check ownself”; the hashtag #ownselfcheckownself emerges on twitter; and gif memes of singh mouthing “ownself check ownself” proliferate singaporean web communities. as online users share these memes and videos, they both allude to the hilarity of “ownself check ownself” and declare a lack of support for the pap’s brand of politics. “ownself check ownself” becomes entextualized into these newer contexts-of-use (silverstein 2011) and becomes a part of the online singaporean lexicon. 1 the hybrid nature of singlish’s makeup (the singlish lexicon is from among others, english, mandarin chinese, southern min languages and malay) and the indefiniteness of whether elements in singlish are code-switches, borrowings, or part of the language itself have resulted in a "less than straightforward" definition of what singlish is (leimgruber 2014); see hall and nilep (2015) for a general review of hybrid linguistic processes. i will explore this in more detail further on in the paper. 3 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 at this point, i will bring up several observations and points of contention. first, there is the rise of the "newly competitive" opposition party and opposition politician that seems to benefit from non-traditional spaces of communication (ortmann 2011) as they use humor to attack and question the ruling legitimacy of the pap. second, we see the paralleled emergence of a historically politically-repressed public on new platforms of participatory media, an electorate who in the past decades, have had few public channels on which to openly congregate. third, we can also identify the centrality of language in this picture, and how ideologies surrounding singlish (and standard singapore english) can reflect, and be a means to, shifting relations of power and social inequality in singapore. by analyzing satirical online discourse of the 2015 general election, i explore the enregisterment (agha 2005) of singlish through the ideological linking of language with non-eliteness and political resistance as singlish-users grapple with questions of linguistic ownership and expertise. as gal (1989:348) observes, the control of representations of reality not only signals the source of social power, but can also reveal loci of conflict. this paper uncovers the various struggles for symbolic domination through the negotiation of what is linguistically real. i start with a summary of the history of politics in singapore, and its current political climate. i examine the legislative changes that have led to the rapid mediatization of singapore politics in the past decade. the change over time in symbiotic relationships between the mainstream media, political organizations, and the electorate in singapore can be understood as the foundation for ideological shifts both politically and linguistically. i then look at data from political rally speeches as well as responses and comments from online users on social media platforms, which indicate the emergence of an alternative linguistic marketplace (bourdieu 1977) in which singlish, in particular, singlish humor, become enregistered as a valued symbol of populist resistance. i discuss how these metadiscursive practices constrain what can be and who can use singlish, as the label becomes an ideological resource through which “anti-pap” can be a part of the construction of singaporean-ness. in order to do that, i employ michael silverstein’s 2003 framework on indexical order, to track how over time, first-order indexicality with correlations to nationality and place can give rise to second-order indexicalities of humor, non-eliteness and political resistance in which a variety gets enregistered as a stabilized set of linguistic features (johnstone, andrus & danielson 2006). i conclude with thoughts on how the nation-state is aligned with certain ways of 4 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 using language, and how speakers complexify the new spaces that they create as they attempt to effect social change. 2. mediatized politics in singapore an analysis of communicative acts and political discourse in singapore requires an understanding of the contemporary political climate in the country. as mentioned in the section above, the people’s action party (pap) have won every election in singapore since 1959, when the colony was first granted self-governance from the british, and have held all or most of the seats in every general election since 1968. throughout the years, the pap have been the political elite in the country, ruling with virtually no opposition. for many singaporeans, the pap name is synonymous with government (agence france-presse 2001). yet a common accusation of the pap in recent years is their growing disconnect with citizens (tan 2012), among other things, singapore’s ministers earning the highest ministerial salaries in the world. under pap’s governance, singapore’s gross domestic product grew steadily from $400 (usd) per capita in 1960 to $61,000 (usd) in 2013 (risse 2014). this was a story endorsed by late founding father lee kuan yew as a “third world to first world” miracle (lee 2012), and singaporeans easily accepted the pap’s “economic pragmatism” as a common sense “non-ideology” (chua 1995) that has since percolated through all aspects of singaporean life. with the ruling pap associated with long term economic stability and fiscal success, opposition parties are left with a very narrow field in which to articulate alternative visions of leading the nation (ortmann 2010). as such, political opposition in singapore have historically been ineffective in having any say in parliament (mutalib 2002). in 2015, the pap won almost 70% of the popular vote, an increase of 9% from the previous election, and 93% of the seats in parliament, ensuring that singapore would be led by a pap supermajority for another five years. 2.1. a climate of fear after the general election results were announced, singapore democratic party leader chee soon juan, published an opinion piece on his personal website, titled "fear and our future" (chee 2015), as he responded to pap’s big win. chee attributed the loss of opposition votes to voters’ "fear[s] at the various stages and aspects of the electoral process", from a fear that the government could somehow figure out who they voted for, or that there might be a “freak election result” and pap would no longer be in power (see khalik & tham 2015). singaporeans voted for the pap 5 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 because the pap, over the years, has "stoke[d] this self-limiting and counter-productive fear", he claims (chee 2015). chee is a prominent opposition party member in singapore having joined singapore politics in 1992. over the years, he has been declared bankrupt after losing defamation suits brought about by former pap prime ministers, and arrested and imprisoned over offenses related to public speaking and illegal assembly (han 2012). chee’s story is as an example to many singaporeans, a contributing factor to the feeling of fear that joining opposition parties might invite retaliation from the government (ortmann 2010). novelist catherine lim refers to the pap government’s history of "responding severely to any criticism of government style of competence" as creating a fear that "silence[s] existing dissident voices and discourage[s] potential ones", omnipresent enough to affect the everyday lives of singaporeans and create self-censorship in their behavior (c. lim 2007). the pap government have been generally cautious about what they have called the irresponsibility of a "free-for-all internet campaigning environment without rules" (british broadcasting corporation 2001). during the 2006 general elections, the media development authority, a statutory board of the singapore government, put out a notice that individuals who use their websites and blogsites to “persistently propagate, promote or circulate political issues relating to singapore”, had to register as holders of political websites (chia, low & luo 2006). the government insisted that this was merely administrative and not “punitive in nature”, a measure to ensure that individuals would not be able to “hide behind the anonymity afforded by the internet in order to influence the electoral process” (bhavani 2006). according to balaji sadasivan, a senior minister of state, individuals needed to realize that "they are still governed by the laws of the land [...] includ[ing] libel" (the straits times 2006). opposition party members reacted strongly to the legislation, concerned that it will affect their campaign strategies and more importantly, shut down political discourse when voters fear breaking the law (chia et al. 2006). such subjugation of the media is not new nor unusual to singaporeans (seow 1998); with media laws that dictate the freedom of the press and protect the pap’s interests over the years, the pap’s ability to shut down contrasting views have led their regime to be described as authoritarian by political observers (rodan 2004; tan 2012). 6 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 2.2. rapid mediatization and a changing climate in 2011, however, amidst social pressure to “keep up with the times”, the pap government amended the parliamentary elections act, cautiously lifting the total ban on online campaigning. since then, singapore politics have undergone a rapid mediatization, as a new political public sphere has emerged from the interaction between traditional print and broadcast media with new platforms of digital and social media. there has been an explosion in the number of websites focused on sociopolitical commentary and news in singapore (salleh 2015), and an increasing social media platform presence: twitter and facebook have grown into spaces for the dissemination of information about political rallies, and youtube and other video streaming sites have provided the platform for election rallies to be edited and broadcast over the internet. for opposition parties, the easing of restrictions governing online content has opened up new platforms to rally for support. their messages can now directly reach the public, unconstrained by the curation of government-controlled traditional media outlets. in fact, the proliferation of online coverage has also pressured the mainstream media to produce “more balanced” coverage, something that was seen as lacking in previous elections. in an online debate organized by nowdefunct political website inconvenient questions, ramesh subbaraman, a former senior reporter with state-owned media organization mediacorp, revealed how television reporters in the past, for example, were instructed to "take close-up shots (so as not) to show [...] large crowds at certain rallies [...] [b]ut that is all gone today" (t. h. tan 2015). these new methods of circulation and recirculation of political messages have also lowered barriers to entry to participating in politics for many singaporeans (mydans 2011), a significant development when compared to the threat of regulation and arrest posed in the previous elections. the participatory structures present in the framework between politician and electorate are unique, especially when interactions play out in a virtual space like the internet. here, i find it useful to refer to erving goffman’s (1981) participant framework in interaction: as political messages circulate across digital domains, the recipients of the message (voters) are separated in space and time from the principal/animators of the message. thus, these voters are not ’ratified listeners’ in the traditional sense, due to their inability to participate discursively at the point when the message is produced. ian hutchby (2006) categorizes this particular participant role as “distributed recipient”, much in the same vein as we might categorize audience members in a broadcast television show. the change in the parliamentary elections act not only allowed for singaporeans 7 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 to enter the conversation without the threat of government oversight or regulation, but presented ways in which these distributed recipients could produce feedback (chovanec & dynel 2015), or become principal-animators in their own right through facebook, twitter or comments on forums and blogs. the once-passive recipient role has shifted to an active one that is afforded production privileges, and it is in these new spaces where politicians, opposition or not, can be under closer scrutiny of the electorate. thus, to understand singapore’s political climate is to make sense of the interdependencies between traditional and participatory media, political organizations, and singaporean citizens, singapore’s form of “mediatized democracy” (mazzoleni & schulz 1999). 3. singlish and the indexical model perhaps the primary question that backgrounds research on singlish in singapore, is how to define what it is and what it does for speakers. the body of work on singlish is substantial, yet its linguistic status has confounded academic commentators. recent structural work on singlish focused on linguistic features often examines the variety’s emblematic sentence-final particles (see wee 2004; l. lim 2007; and hiramoto 2012 for examples) or tense, aspect, and modality markers (bao 2005; nomoto & lee 2012). for example, bao zhiming’s 2005 work on the singlish use of already in clause-final positions investigates what he sees as chinese influence, concluding that its coding of aspectual meanings echoes that of the chinese particle le. hiroki nomoto and nala lee (2012) argue that the use of got in singlish is neither a tense nor aspect marker, as previously classified, but rather a realis modality marker that codes temporality through the requirement of the factual status of situations at the time of utterance. while the value of structural studies cannot be understated, most of such work on singlish identify specific iconic elements whose “singlishness” will never be refuted. it is not surprising, then, that much of the academic research on singlish is social in approach. this research, influenced by newer sociolinguistic research paradigms (cf. bucholtz & hall 2005), involves extensive studies on the role of singlish as an ideological resource for identity construction, analyzing its use within a variety of political, economic, and cultural contexts (alsagoff 2010). as a result, just in the past decade, the sociolinguistic body of research that models singlish in relation to an overtly prestigious singapore standard english has been extensive and detailed (e.g., alsagoff 2010, 2013; wee 2011; chua 2011; leimgruber 2013). in these studies, the movements between the standard variety and singlish are seen as fluid, dynamic 8 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 achievements, as speakers negotiate different stances and identities in using and talking about singlish. this approach implies that for many of these researchers, a code-switch from singlish to english is an ideological one, and the overall effect of this mixing allows singaporeans to construct equally dynamic intersubjectivities. lubna alsagoff’s (2010) cultural orientation model, for example, explores glocalization as an explanation to the lack of any clear boundaries between the two languages. she focuses on the simultaneous global and local contextual pressures that speakers must negotiate as they interact, and while speakers do seem to orient to one variety or another as they produce speech, the overall effect is one of mix. the argument here is that this unbounded mixture between the two varieties is itself important to understanding the ideological tensions that speakers face. jakob leimgruber (2012) has proposed that singlish can be best described by an indexical model, drawing on the frameworks of indexical order (silverstein 2003) and indexical field (eckert 2008). in a variety like singlish, where there is no “easily identifiable matrix language” (leimgruber 2012:11), he argues that it is perhaps more productive to think of singaporean discourse as elements (or features) associated with the multiple codes used in the region. by relegating the correlation between language variety and linguistic form to a metapragmatic and possibly, metadiscursive level, we can then investigate the social meanings of singlish tokens without taking away or assuming that speakers themselves are orienting to absolute “bounded entities” (blommaert & rampton 2011) as they interact (leimgruber 2012). i take leimgruber’s insightful work as the starting point of my inquiry into the process of enregisterment of singlish. i argue that the legitimization of singlish is produced through the more “authentic” (cf. woolard 2016) authority, the singaporean electorate. 3.1. social meanings of singlish the concept of orders of indexicality, as proposed by michael silverstein in 2003, is interested in the process by which linguistic forms acquire social meaning; how indexicalities are born through social interaction. leimgruber’s work on singlish, while useful in revealing the circularity of a variety-first approach (speakers, to be more singaporean, use singlish, therefore forms are singlish are when speakers “do being singaporean”), does not specifically address what social meanings are increasingly linked to singlish tokens, or how the selection of what constitutes singlish is done by speakers. if there is no “easily identifiable matrix” in singlish, we perhaps need to turn our attention to the processes at the second order of indexicality, in which 9 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 enregisterment of a variety can occur, and what features get stabilized in the process (johnstone et al. 2006). we first consider singlish at its first order of indexicality. according to silverstein, nth-order indexicals are features whose accounts of their correlation with use are "scientific" (2003:205), a link that can be easily identified by cultural outsiders. in the case of singlish features, we see this presented rather evidently in the indexical link between singlish use and “singaporean-ness”: a “local”-ness that is distinctively different from a “global” orientation (alsagoff 2010). singlish is what is used by singaporeans, what all singaporeans know, and to use singlish is to be singaporean. at this level of indexicality, any conclusion on why singaporeans use singlish would therefore be analytically unproductive; it does not provide any further insight into where the boundaries of singlish are with regards to its co-existence with standard singapore english, that being one of the aims of researchers working on defining singlish. this indexical phenomenon, then, becomes productive when it gains enough saliency to be “available for social work” (johnstone et al. 2006:82), where the nth-order correlation can give rise to a correlation at the level of n+1. this n+1th-order indexicality is driven by ethno-pragmatic interpretations of nth-level indexicality, and is formed by ideologies surrounding who uses singlish. it is at this level that my interest in singlish lies. the following sections will investigate, in more detail, the processes in which singlish tokens gain particular social meanings. 3.2. humor, political resistance, and non-eliteness one of the earliest and most prominent examples of political satire occurred as a direct result of the 2006 debate on a new law that prohibited proselytizing on the internet. many singaporeans were confused by the ambiguity of the government’s position: any online website that were found to "persistently propagate, promote or circulate political issues relating to singapore" (chia et al. 2006) were “political”, and therefore, needed to be registered with the media development authority of singapore. to get around the legislation, two of singapore’s most prolific bloggers and podcasters, lee kin mun (mrbrown) and benjamin lee (mr miyagi), started the “persistently non-political podcast” during the 2006 general elections as a way to report on election happenings using analogous, fictional stories that appealed to local knowledge. unable to overtly “circulate political issues”, they spoofed a particular incident where the pap was seen to be haranguing an opposition party member for a lie he had told, analogizing it as an exchange between a hawker 10 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 vendor selling bak chor mee, popular national dish of minced pork noodles, and his customer. a portion of the podcast is reproduced below: (2) customer: uh.. wait wait, hang on. this has tur kua in it. wait, wait, hang on. this has pork liver in it. hawker: yah lah, is got tur kua, liver one what. yes, (this dish usually) has pork liver. customer: but i said i didn’t want tur kua. hawker no you didn’t. customer: yes i did. [...][several lines redacted] hawker: no you didn’t and i can prove it to you ah. customer: very well, prove it! customer: the what? hawker: nah, you see, you point to the mee pok, then, you say dry. then, you point to the chilli, then, you shake your head. you never say you don’t want to have the tur kua! here, you see, you pointed to the flat noodles then you said dry noodles (not in soup). then, you pointed to the chilli, then you shook your head. you never said you didn’t want the pork liver! the hawker vendor in the above example represents the pap, and the customer the hapless opposition party member. this was a spoof of an incident where the pap was seen to be continuously demanding an explanation from an opposition party member who had apologized for his mistake, inadvertently also making viral the singlish phrase “sorry also must explain!”. this event was generally interpreted by singaporeans as an inconsequential and petty act by the ruling pap, something further propagated by the release of this podcast. this episode of the “persistently non-political podcast” went viral and was shared more than 100,000 times (koh 2006) it propelled mrbrown to internet celebrity status, and gained enough attention that pap prime minister lee hsien loong, in his national day rally speech later that year, referenced the podcast in a bid to demonstrate his government’s willingness to adapt to new approaches in reaching out to singaporeans. the success of that podcast episode led mrbrown to produce more episodes focusing on politics in singapore, some of them featuring the same hawker vendor and his customer. the show also released several free downloadable desktop wallpapers on its website, featuring an anthropomorphized “tur kua” character raising a fist in the air in anger (see figure 2 below), together with a particularly popular phrase from the episode, "why you say you tell me you dowan tur kua when you didn’t say you dowan tur kua?" ’why did you say (that) 11 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 you told me (that) you didn’t want pork liver when you didn’t say (that) you didn’t want pork liver?’ figure 2. tur kua desktop wallpaper available for download on mrbrownshow.com what is fascinating in this example is the how mrbrown and mr miyagi have exploited linguistic signs in one satirical podcast, and the specific forms that then get entextualized in new contexts like the desktop wallpaper. in the podcast episode, the original incident was taken from its political, “formal” context and supplanted in a distinctively local setting, the hawker center. the linguistic tokens used here also follows this backdrop, with the hawker vendor speaking in a local, non-standard register, notably different from the customer, who continues the conversation in a recognizably more “standard” variety. in the desktop wallpaper, the phrase that is seen as most salient, thus worthy of enshrining, is also an utterance made by the hawker vendor. as a form of ridicule, the podcast succeeded as one of the first examples of political resistance on the internet” it achieves both a very funny, native re-frame of the pap’s actions, as well as simultaneously mocks the draconian law against putting political content online. most importantly, the linguistic signs employed here are “very singaporean”, which then opens up the potential for their categorization as singlish. ridicule here acts as a form of rebellion (billig 2005), where humor can be found in both 1) the challenging of authority, as well as 2) in “breaking the codes of 12 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 language”, expected rules of communication when talking about politics (2005:207). by using very recognizable, local contexts, mrbrown and mr miyagi draw social boundaries of “us” versus “them”, and bring singaporeans who see themselves as non-elite and anti-establishment together through the mockery of the elite class. 3.3. the opposition politician opposition political parties, due to decades of suppression by the pap, are a natural symbol of political resistance to most singaporeans. the main focus when developing the brand of an opposition candidate has always been to present oneself as an alternative voice to an elite pap who have been in power for so long, that they are out of touch with singaporeans on the ground (the workers’ party 2015). “non-eliteness” and “anti-pap” are second-order indexicalities that have been increasingly linked to non-standard features, and thus it is easy to see how they might be taken up as a resource in opposition politician speech. recall at this point the example of “ownself check ownself” at the beginning of this paper. pritam singh cannot just give his entire speech in the local register. english, or standard singapore english, is the language of the government and administration, the main medium of instruction in schools, and was crucial to singapore’s ability to participate in the global marketplace (goh & gopinathan 2008). english language proficiency has always been seen as integral to the singaporean brand, commodified and celebrated as skill (heller 2010). political candidates like singh, in their speeches, need to be able to manage conventionalized expectations of using english, as it carries indexicalities of personal and academic success. when promoting the worker’s party manifesto throughout the rest of his speech, singh speaks entirely in english, the overtly prestigious language. at the point where he makes a strong criticism of the pap, he asks the crowd, “is this the future we want for singapore or our children in the next fifty years?”, inciting a shouted response “no!”. he then follows up with “ownself check ownself!” if we take singlish to be an indeterminate nebula of signs, then the question we should be asking is how, then, does “ownself get ownself” get enregistered as singlish? the following section examines how the phrase acquires legitimacy at the n+1th order of indexicality. as mentioned, singh deviates from a speech given almost completely in standard singapore english. “ownself check ownself” is not recognizable as english, yet this phrase is creative, never heard before in public settings, and (possibly) uncommon in day-to-day interaction, which makes it hard to categorize definitively as singlish either. however, it is used in the context of “politically 13 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 resisting”, as a criticism of the ruling party, at a political rally held by the opposition. it incites laughter from the audience, an example of one more non-standard expressions used to mock authority. singh, like mrbrown and mr miyagi before him, in a jocular moment, "broke[...] the codes of language" (billig 2005), the political register expected of all singaporeans running for office, again drawing the lines between “us” and “them”, “elite” and “non-elite”, “standard” and “non-standard” using a complex of multiple signs and exploiting their potential indexicalities. most importantly, his words can be branded as singlish in metapragmatic discourse surrounding his speaking abilities. an analysis of comments of online users further reveals how singh’s speech style is taken up by the broader public and linked to his identity as an opposition politician. although his speech was peppered with singlish words, he came out as being intelligent, persuasive, and above all, eloquent. he has been hailed as the obama of singapore by some people on the internet. reddotrevolver, online opinion article, 3 may 20112 having listened to his rally speeches, you are right ps will give the pap dogs problems when he is elected. he is indeed an eloquent speaker. sharp and witty too. user golden dragon, sammyboyforum3 here, we see that singh is hailed as eloquent, persuasive, and witty, and a potential formidable foe in parliament against the pap. in the first comment, the commentator reddotrevolver highlights singh’s use of singlish, glosses over its non-standard connotations (“although his speech was peppered with singlish words”), and characterizes him as a great speaker. for this online user, the presence of a non-standard variety in singh’s speech does not affect the way he “sounds” in a negative way, in fact according to user golden dragon, he is eloquent and witty. note too that 2 reddotrevolver (2011, may 3). pritam singh roasts the pap. online: http://asiancorrespondent.com/2011/05/pritam-singhroasts-the-pap/; accessed october 4, 2015. 3 golden dragon (2011, may 2). must see: wp pritam singh very powerful speech for voters all over singapore. sammyboyforum. [forum comment]. online: http://www.sammyboy.com/archive/index.php/t91083.html?s=b76ac2ae0f4d2923d0bdfae091ff897b; accessed october 5, 2015. 14 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 there is no further explanation of which words were singlish, just a general sense that singh’s speech was “peppered” with them. 3.4. legitimizing singlish through entextualization days after singh’s speech at the rally, known anti-establishment sociopolitical website all singapore stuff posted a video on their facebook page (see figure 3 below). figure 3. screenshot of video posted on facebook by all singapore stuff it comes with the descriptor, "watch lee hsien loong and pritam singh argue over ownself check ownself" and the hashtags #funny, #ownselfcheckownself, #ge2015 and #eastcoastwillfall. the 40-second video injects edited clips of singh shouting “ownself check ownself” with that of pap prime minister lee’s own rally speech segments, ultimately resulting in a video that mocks the pap’s position on “ownself” checking and balancing. this video has been viewed over 98,000 times and shared 1,472 times as of 1 july 2017. we see the lines very clearly drawn between the opposition and the ruling politician, with abrupt jump cuts from one to the other, the wp pitted against the pap. this was a clip meant to produce laughter (as evidenced 15 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 by #funny) and chooses “ownself check ownself” as its main inspiration. figure 4 shows other examples of the extextualizations of the phrase on twitter, where the hashtag #ownselfcheckownself started being used on twitter in the days following singh’s speech. figure 4. screenshots of tweets that utilize the hashtag #ownselfcheckownself the first tweet by user axel (@axeltann) shows a thinly veiled criticism of the pap, and the activation of the hashtag #ownselfcheckownself marks user axel as a knowing member of the public against the pap, someone who both recognizes “ownself check ownself” and wants to make a stand about the ruling party. in the second tweet, we see user shin sheng (@lupcheongster) using the hashtag in a completely different political context: referring to the 2015 southeast asian haze crisis, during which there was severe air pollution in neighboring southeast asian countries over a period of several months caused by illegally-created indonesian forest fires which spread quickly in the dry season. with haze blanketing parts of indonesia, malaysia, and singapore, the singapore government’s initial offers of help to combat the fires were declined by the indonesian state, who stated that they have sufficient resources to deal with the crisis (a. tan 2015). by posting “don’t worry singapore, indonesia can #ownselfcheckownself”, user shin sheng entextualizes “ownself check ownself” in a new context-of-use, further legitimizing the expression as a bona fide one that has pragmatic functions outside of its original context. the tweets, shares, likes, and new uses of “ownself check ownself” in social media posts function as "identity statements expressing, pragmatically and metapragmatically, membership of some groups [...] not held together by high levels of awareness and knowledge of deeply shared values and functions, [...] but by loose bonds of shared, even if superficial interest" (varis & 16 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 blommaert 2015:35). by sharing and liking the post, or activating a hashtag, groups are brought together on demand, as individual users select their participation, indicating their recognition of the original context-of-use and then showing appreciation for its adaptation into a different modality. humor is used in these ways online to bring groups together, not only because they think something is funny, but because what is funny is only funny if you have similar viewpoints about the pap government. the affordances of these platforms allow an otherwise-invisible audience to rise up against a perceived more-powerful group, and the opening up of the internet has given the singapore electorate spaces they previously never had to exercise judgment on their government, through “local” resources like singlish. in this process, we can also identify another what the singaporean electorate find is particularly worthy-of-comment. “ownself check ownself” is what is ruminated on (cf. silverstein 2011), over and over again. online commenters recognize the semiotic productivity of the phrase, and in their entextualization of it in newer contexts. they draw their own boundaries around language through the way they discuss the expression in their own posts, legitimizing “ownself check ownself” as an emblematic expression of singlish. literary critic gwee li sui, in a may 2016 opinion piece about politics and singlish in the international edition of the new york times, pointed out that “even politicians and officials are using it”, and specifically identifies singh’s “ownself check ownself” at the workers’ party rally as an instance of singlish use. the expression is a new, creative, and now-public sign that has been increasingly fossilized through each new metapragmatic re-use and entextualization. it becomes singlish, as it accumulates citations through the codifying of resistance against the pap elite. 4. discussion and conclusion in the short paper above, i have outlined, in as much detail as i can given the limitations on space, the processes through which singlish terms are enregistered as part of a language variety. the enregisterment process, as linguistic anthropologists have suggested across multiple publications, is highly ideological. i have illustrated how this process depends on the ways that subjects experience their sociopolitical environment. singlish, as a language, undergoes constant shifts, and this paper hopes to provide some clarity into these processes of enregisterment, focusing on the instantiations of speaker metapragmatic use which can lead to the emergence of n+1th orders of indexicality. the shifting interrelationships between traditional and new media, political 17 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 actors and political observers has shown us how language can change, both reacting or being reacted to, as political agents conceptualize its ideological boundaries. lastly, johnstone et al. (2006) reminds us that orders of indexicality are in dialectical relationships with each other, and thus we see why singlish can both be the resource and outcome of ongoing singlish enregisterment processes. second-order indexicalities (non-eliteness, political resistance through humor) can in turn inform first-order realities, and affect what linguistic features can picked up as singlish as singaporeans become more and more aware of the value of singlish as an indicator of their “groupness”. the fight to gain social power by a historically silenced electorate will thus almost definitely involve symbols of localness and non-eliteness, and singlish, with each citation, iconizes more and more to perform that duty of resisting the powerful. however, we see that these forms of resistance only showing up in instances of rebellious humor, of satire and parody. singh, who as an opposition politician embodies political rebellion in singapore, can only pick at specific moments in his speeches to use singlish. he, together with the electorate, is constrained by societal expectations of what language is appropriate for which context. the change in attitudes towards singlish, however, cannot be ignored, and it is in the political realm where we see it gaining the most traction in public settings. we cannot forget, however, to be skeptical of such developments. progressive change from past practices might look like a step in the right direction, yet attention needs to be paid to what current work they are achieving. the almost military precision with which such humor is done, constrained into very specific settings, is the current domain of singlish. according to billig (2005), rebellious humor does not necessarily have rebellious effects, and as speakers reveal themselves as “captive” to the demands of having to incorporate humor as they make a stance against the pap, the overall effect could be one that strengthens instead of subverting power. singlish becomes synonymous with humor, with non-eliteness, and the language is further delegitimized as a variety in relation to english. in a similar vein, the pap is increasingly associated with english, and benefits from its indexical links to global success and professionalism. further investigation how political actors navigate the singlish-english indexical field needs to be done to uncover the dynamics of language and power in the singapore context. this would also provide new insight into the role singlish plays in singaporean society, and in the construction of singaporean identities. 18 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 references agence france-presse. 2001. pap synonymous with singapore government. november 4, 2001. retrieved may 8, 2014 from http://www.singapore-window.org/sw01/011104a2.htm agha, asif. 2005. voice, footing, enregisterment. journal of linguistic anthropology 15(1).3859. alsagoff, lubna. 2010. english in singapore: culture, capital and identity in linguistic variation. world englishes, 29(3).336–348. retrieved 2016-04-19, from http://onlinelibrary.wiley.com/doi/10.1111/ j.1467-971x.2010.01658.x/abstract alsagoff, lubna. 2013. singlish in hybridity: the dialogic of outer-circle teacher identities. the global-local interface and hybridity: exploring language and identity, ed. by rani rubdy and lubna alsagoff, 265-281. multilingual matters | channel view publications. bao, zhiming. 2005. the aspectual system of singapore english and the systemic substratist explanation. journal of linguistics 41(2).237-267. bhavani, k. 2006, april 20. election advertising is prohibited on the net. the straits times forum. retrieved from https://www.mci.gov.sg/pressroom/news-andstories/pressroom/2008/1/election-advertising-is-prohibited-on-the-net?page=159 billig, michael. 2005. laughter and ridicule: towards a social critique of laughter. london; thousand oaks: sage. blommaert, jan and rampton, ben. 2011. superdiversity. diversities, 13(2).4-18, from http://www.mmg.mpg.de/fileadmin/ user_upload/subsites/diversities/journals_2011/ 2011_13-02_gesamt_web.pdf bourdieu, pierre. 1977. the economics of linguistic exchanges. social science information, 16.645–668. doi: 10.1177/ 053901847701600601 bucholtz, mary and hall, kira. 2005. identity and interaction: a sociocultural linguistic approach. discourse studies 7(4-5).584-614. british broadcasting corporation. 2001, august 14. singapore net law dismays opposition. bbc news. retrieved from http://news.bbc.co.uk/2/hi/asia-pacific/1490425.stm chee, soon juan. 2015. fear and our future. cheesoonjuan.com. retrieved from http://www.cheesoonjuan.com/ home/fear-and-our-future chia, sue-ann; low, aaron; and luo, serene. 2006, april 5. opposition parties slam podcast ban rule. the straits times. retrieved from http://www.international.ucla.edu/asc/article/42107 19 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 chovanec, jan and dynel, marta. 2015. participation in public and social media interactions. john benjamins publishing company. chua, beng huat. 1995. communitarian ideology and democracy in singapore. routledge. chua, catherine siew kheng. 2011. singapore’s e(si)nglish-knowing bilingualism. current issues in language planning 12(2).125–145. cochrane, joe. 2015. singapore vote will test long ruling party’s grip on power. the new york times. retrieved from https://www.nytimes.com/2015/09/10/world/asia/singapore-vote-willtest-long-ruling-partys-grip-on-power.html eckert, penelope. 2008. variation and the indexical field. journal of sociolinguistics, 12(4).453– 476. gal, susan. 1989. language and political economy. annual review of anthropology, 18(1).345367. goffman, erving. 1981. forms of talk. university of pennsylvania press. goh, c. b., and gopinathan, s. (2015). the development of education in singapore since 1965. toward a better future: education and training for economic development in singapore since 1965, ed. by lee sing kong, tan jee peng, birger fredriksen, and goh chor boon, 12–38. washington, dc: the world bank publications and the national institute of education (nie) at nanyang technological university. hall, kira and nilep, chad. 2015. code-switching, globalization, and identity. handbook of discourse analysis, ed. by deborah tannen, heidi e. hamilton and deborah schiffrin, 597619. malden, ma: wiley-blackwell. han, kirsten. 2012. democratically speaking: chee soon juan promotes nonviolent action in singapore. wagingnonviolence.org. retrieved from https://wagingnonviolence.org/feature/democratically-speaking-chee-soon-juan -promotesnonviolent-action-in-singapore/ heller, monica. 2010. the commodification of language. annual review of anthropology, 39(1).101–114. retrieved from https://doi.org/10.1146/annurev.anthro.012809.104951 doi: 10.1146/annurev.anthro.012809.104951 hiramoto, mie. 2012. pragmatics of the sentence-final uses of can in colloquial singapore english. journal of pragmatics 44(6).890-906. 20 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 hutchby, ian. 2006. media talk: conversation analysis and the study of broadcasting. glasgow: open university press. johnstone, barbara; andrus, jennifer; and danielson, andrew. e. 2006. mobility, indexicality, and the enregisterment of "pittsburghese". journal of english linguistics, 34(2).77–104. khalik, salma and tham, yuen-c. 2015. no guarantee pap will be in government after polls: khaw boon wan. the straits times. retrieved from http://www.straitstimes.com/politics/noguarantee-pap-will-be-in-government-after-polls-khaw-boon-wan koh, leslie. 2006. net spoof too funny for serious politics? the straits times. retrieved from http://www.international.ucla.edu/asc/article/47020 lee, kuan yew. 2012. from third world to first: the singapore story, 1965-2000. vol. 2. marshall cavendish international asia pte ltd. leimgruber, jakob. 2012. singapore english: an indexical approach. world englishes, 31(1).114. leimgruber, jakob. 2013. singapore english: structure, variation, and usage. cambridge university press. (studies in english language series). leimgruber, jakob. 2014. singlish as defined by young educated chinese singaporeans. international journal of the sociology of language. retrieved from https://www.degruyter.com/view/j/ijsl.2014.2014 .issue-230/ijsl-2014-0026/ijsl-20140026.xml doi: 10.1515/ijsl-2014-0026 lim, catherine. 2007, october 29. sg daily special: an open letter to the prime minister by catherine lim. singaporedaily.net. retrieved from http://singaporedaily.net/2007/10/29/sgdaily-special-an-open-letter-to-the-prime-minister-by-catherine-lim/ lim, lisa. 2007. mergers and acquisitions: on the ages and origins of singapore english particles. world englishes 26(4).446-473. mazzoleni, gianpietro and schulz, winfried. 1999. "mediatization" of politics: a challenge for democracy? political communication, 16(3).247–261. retrieved from http://www.tandfonline.com/doi/abs/10 .1080/105846099198613 doi: 10.1080/105846099198613 mutalib, hussin. 2002. constitutional-electoral reforms and politics in singapore. legislative studies quarterly, 27(4).659-672. 21 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 mydans, seth. 2011. opposition makes inroads in singapore. the new york times. retrieved from http://www.nytimes.com/2011/05/08/world/asia/08singapore.html nomoto, hiroshi and lee, nala h. 2012. realis, factuality and derived-level statives: perspectives from the analysis of singlish got. building a bridge between linguistic communities of the old and the new world: current research in tense, aspect, mood and modality 25.219-239. ortmann, stephan. (2010). politics and change in singapore and hong kong: containing contention. routledge. ortmann, stephan. 2011. singapore: authoritarian but newly competitive. journal of democracy, 22(4).153–164. retrieved from https://muse.jhu.edu/article/454067/summary philomin, laura e. 2015. check and balance a seductive lie: esm goh. todayonline. retrieved from http://www.todayonline.com/singapore/check-and-balance-seductive-lie-esmgoh risse, mathias. 2014. from third world to first what’s next? singapore’s obligations to the rest of the world from a human rights perspective. hks working paper no. rwp14007.1–25. retrieved from https://ssrn.com/abstract=2410434 rodan, garry. 2004. transparency and authoritarian rule in southeast asia: singapore and malaysia. routledge. salleh, nur asyiqin mohamad. 2015. political websites creating a buzz in singapore. the straits times. retrieved from http://www.straitstimes.com/singapore/political-websites-creating-abuzz-in-singapore. seow, francis t. 1998. the media enthralled: singapore revisited. lynne rienner publishers. silverstein, michael. 2003. indexical order and the dialectics of sociolinguistic life. language and communication, 23(3).193-229. silverstein, michael. 2011. presidential ethno-blooperology: performance misfires in the business of" message"-ing. anthropological quarterly, 84(1).165–186. retrieved from https://muse.jhu.edu/article/417394/summary silverstein, michael and urban, greg. 1996. natural histories of discourse. university of chicago press. 22 colorado research in linguistics, vol. 24 [2019] https://scholar.colorado.edu/cril/vol24/iss1/5 doi: http://dx.doi.org/10.33011/cril.24.1.5 tan, audrey. 2015. jakarta again declines singapore’s help to fight haze. the straits times. retrieved from http://www.straitstimes.com/singapore/environment/jakarta-again-declinessingapores-help-to-fight-haze tan, kenneth p. 2012. the ideology of pragmatism: neo-liberal globalisation and political authoritarianism in singapore. journal of contemporary asia 42(1).67-92. tan, netina. 2014. the 2011 general and presidential elections in singapore. electoral studies 35.374–378. tan, tarn how. 2015. ge 2015: media coverage more “balanced”, say ex-insiders. the online citizen. retrieved from https://www.theonlinecitizen.com/2015/09/11/ge-2015-mediacoverage-more-balanced-say-ex-insiders/ the straits times. 2006, april 4. political podcasts, videocasts not allowed during election. the straits times. retrieved from http://international.ucla.edu/institute/article/42036 the worker’s party. 2015. manifesto 2015: enpowering your future. worker’s party website. retrieved from http://www.wp.sg/manifesto/ varis, piia and blommaert, jan. 2015. conviviality and collectives on social media: virality, memes, and new social structures. multilingual margins 2(1).31-45. wee, lionel. 2004. reduplication and discourse particles. singapore english: a grammatical description, ed. by lisa lim, 105-126. amsterdam/philadelphia: john benjamins. wee, lionel. 2011. metadiscursive convergence in the singlish debate. language and communication 31(1).75–85. woolard, kathryn a. 2016. singular and plural: ideologies of linguistic authority in 21st century catalonia. oxford studies in anthropology of language. oxford university press. 23 khoo: the role of singlish humor in the rise of the opposition politician in singapore published by cu scholar, 2019 colorado research in linguistics 6-2019 “ownself check ownself”: the role of singlish humor in the rise of the opposition politician in singapore velda khoo recommended citation microsoft word khoo-cril2019-final.docx microsoft word bonial-cril2021-proof_final.docx 1 précis of take a look at this! form, function, and productivity of english light verb constructions claire bonial u.s. army devcom army research laboratory english light verb constructions (lvcs), such as make an offer and take a bath, are semi-productive constructions: while some novel combinations of a light verb and eventive or stative noun are acceptable, others are not. lvcs tend to occur in families defined by a shared light verb and semantically similar nominal complements, but it remains mysterious precisely how such families are circumscribed; this presents both theoretical and natural language processing (nlp) challenges. this research first aims to address the void in linguistic resources identifying lvcs in a consistent fashion with the development of annotation guidelines for lvcs within the propbank project. using the resulting annotated corpus of lvcs, another theoretically important question of why lvcs exist alongside semantically similar lexical verbs is addressed: corpus evidence demonstrates that the ease and variety with which lvcs can be modified is the primary motivating factor for their use over a lexical verb. finally, large-scale acceptability studies are used to examine the constraints on lvc productivity; the results reveal the importance of statistical preemption in modeling productivity. keywords: light verbs, constructions, semantic roles, productivity, natural language processing 1. introduction a key question in linguistics and cognitive science is how people learn and apply the seemingly idiosyncratic constraints of a language. for example, in english, one can tell me the facts but cannot *explain me the facts (goldberg, 2011).1 similarly, one can have a drink, but cannot *have an eat (wierzbicka, 1982). how do speakers know this? from a usage-based, construction grammar perspective (e.g., goldberg, 1995, 2006), the crux of this issue is why certain constructions (pairings of form and meaning) are compatible, or incompatible, with certain lexical items. specifically of interest for this research, why is drink compatible within the have light verb construction (lvc) (jespersen, 1942) while eat is not? the answer to this question carries ramifications for larger issues of what is stored in the mental lexicon, and how human grammar develops. some argue that construction grammar makes certain claims on the storage and processing of lexical items: phrasal constructions, as pairings of form and meaning, are stored in the mental lexicon in the same way that individual lexical items colorado research in linguistics, volume 25 (2021) 2 are stored in the lexicon. this suggests that constructions are not decomposed or analyzed compositionally – according to the semantics of each individual lexical item (piñango, mack & jackendoff, 2006; wittenberg & piñango, 2011). a related perspective from another usage-based approach, emergent grammar (e.g., hopper, 1998), predicts that constructions are extended by semantic analogy to an existing, high-frequency exemplar construction. the validity of these views is explored here with respect to lvcs. the questions surrounding lvc acceptability also present challenges for natural language processing (nlp) systems. in general, nlp systems rely on a combination of training data, which captures usage patterns of lexical items, as well as computer-readable lexicons, which capture some level of the meaning of a word. complete coverage of lvcs in a lexicon is made impossible by the fact that lvcs are semi-productive: speakers can extend a construction’s template creatively with novel combinations of lexical items. however, that productivity is constrained such that not all combinations are acceptable. as a result, we simply cannot list every lvc in a lexicon. furthermore, novel, creative usages will be quite rare in training data, and will be greatly outnumbered by more conventional usages of the same verbs, precluding fully unsupervised approaches to this problem. nevertheless, we can supplement nlp systems with knowledge from linguistics and psycholinguistics in order to model patterns of productivity and better estimate likelihoods that a given combination will be acceptable. 2. summary a thorough treatment of these issues has been fettered by a lack of resources identifying lvcs in a consistent fashion. thus, this research first aims to address the void in linguistic resources with the development of annotation guidelines for lvcs within the propbank project (palmer et al., 2005) (§4). using the resulting corpus of lvcs, another theoretically important question of why lvcs exist alongside semantically similar lexical verbs is addressed: corpus evidence demonstrates that the ease and variety with which lvcs can be modified is the primary motivating factor for their use over a lexical verb (§5.1). finally, large-scale acceptability studies are used to examine the constraints on lvc productivity; the results reveal the importance of statistical preemption in modeling productivity (§5.2). the impacts of this research on (psycho)linguistics and nlp are discussed in closing (§6). précis of take a look at this! form, function, and productivity of english lvcs 3 3. background: light verb constructions english lvcs are an ideal case for studying integral issues of grammar as lvcs are semicompositional, semi-productive constructions that tend to have semantically similar lexical verb counterparts (e.g., make an offer vs. offer) – a fact which runs contrary to assumptions in linguistic theories that two competing forms are rarely maintained in a language, unless they serve distinct purposes (grice, 1975). each of these characteristics is explained in more detail below. the theories of construction grammar posit that there is a continuum from wholly fixed, noncompositional idioms like kick the bucket, in which no component part expresses the meaning of die; to semi-compositional constructions like lvcs, in which the component parts do contribute lexical meaning, and there is some flexibility as to what elements can form an lvc; to purely compositional language, in which words combine freely and productively according to syntactic rules. this continuum is illustrated in figure 1. as semi-compositional constructions, lvcs cannot be interpreted in a completely literal fashion (take a seat and take a bath are not interpreted as events where chairs and bathtubs are taken somewhere, but instead as sitting and bathing events). speakers must recognize the unique nature of these constructions and interpret them according to the event semantics denoted by the noun. furthermore, our computational systems must do this: both natural language understanding and machine translation (among other computational tasks) require that lvcs are delineated from other usages of the same verbs and interpreted uniquely. parallel to the continuum of compositionality, there is a continuum of productivity. on one extreme, purely compositional language is thought to be fully productive. on the other, some idiomatic constructions are wholly fixed, allowing for no substitution of lexical items (e.g., *he punted the bucket). lvcs are semi-productive (nickel, 1978): while some novel combinations figure 1: continuum of language compositionality. colorado research in linguistics, volume 25 (2021) 4 of a lv and noun are acceptable, others are not. lvcs tend to occur in families defined by a shared lv and semantically similar nominal complements, but it remains mysterious precisely how such families are circumscribed. for example, a variety of nouns denoting communication events can combine with make: (1) make a speech/declaration/proclamation/announcement yet other semantically similar nouns may not form acceptable combinations: (2) ?make a yell while some research has tried to pinpoint the exact semantic constraints that make a given combination acceptable (e.g., wierzbicka, 1982), other research has focused on the importance of frequency and analogy in extending particular constructions. emergent grammarians (e.g., hopper, 1998; bybee, 2006, 2010) hypothesize that semi-productive constructions are extended by semantic analogy to an existing, high-frequency exemplar construction; therefore, novel constructions that are semantically very similar to high-frequency exemplars are predicted to be acceptable. lvcs in english and romance languages are somewhat unique cross-linguistically because they tend to have semantically similar synthetic verb counterparts (zarco, 1999). for example (from coca (davies, 2008)): (3) she appeared with me on vh1 "celebrity rehab." (4) bahrain’s king hamad made a rare appearance on television. this runs contrary to assumptions in linguistic theories that two competing forms are rarely maintained in a language unless they serve distinct purposes (e.g., grice, 1975). what distinct purposes might these forms serve? why do english lvcs exist alongside counterpart synthetic verbs, especially given that synthetic verbs are arguably the more efficient variant form (zipf, 1949)? précis of take a look at this! form, function, and productivity of english lvcs 5 unfortunately, these questions and the issues surrounding lvc usage and productivity have been difficult to address without a large-scale resource providing a markup of both lvcs and counterpart verbs. thus, this research first undertakes the challenge of developing a resource that provides comprehensive annotations of lvcs. 4. approach: development of the propbank corpus of lvc annotations the goal of propbank is to supply consistent, general-purpose labeling of semantic roles across different syntactic realizations. with over two million words from diverse genres, the annotated corpus supports the training of automatic semantic role labelers, which in turn support other nlp areas, such as summarization and question-answering. propbank annotation consists of two tasks: sense and role annotation. the propbank lexicon provides a listing of the coarse-grained senses of a verb, noun, or adjective relation, and the roles associated with each sense (known as a roleset). the roles are listed as argument numbers corresponding to relation-specific roles. for example: offer-012 arg0: entity offering arg1: commodity arg2: price previous versions of propbank maintained separate rolesets for the verb offer, and the related nouns offer/offering, but, as part of this research, these have been combined to provide parallel annotations of related usages (bonial et al., 2014): (5) [ne electric]arg0 offered [$2 billion]arg2 [to acquire ps]arg1 (6) [ne electric]arg0 made an offer of [$2 billion]arg2 [to acquire ps]arg1 prior to this research (see also hwang et al., 2010), propbank 1.0 had no special guidelines for lvcs. therefore, example 6 would have been annotated according to the semantic roles found in the verb roleset, make-01, which is, at best, metaphorically related to the lvc usage (see 9 for a discussion of annotation with make-01). furthermore, this practice conflated lvc usages like colorado research in linguistics, volume 25 (2021) 6 make an offer with heavy, literal usages of the same verb, such as make a cake, in which make is used in its full, creation sense. lvcs require a unique annotation procedure that represents the event semantics of the construction stemming from the noun, as opposed to the verb. first, however, we must recognize lvcs consistently. this has been problematic, given that definitions of lvcs cross-linguistically and within a language remain nebulous. nonetheless, lvcs in english are generally defined as consisting of a semantically general, highly polysemous verb and a noun denoting an event or state (e.g., butt, 2003). more detailed definitional aspects of lvcs remain debatable. this basic definition does not allow for consistent manual, much less automatic, detection of lvcs since surface-identical forms can be lvcs or heavy usages of the same verbs: (7) he took a drink of the soda. (lvc) = he drank the soda. (8) he took a drink off the bar. (non-lvc) ≠ he drank the bar. such usages are indistinguishable with respect to their syntactic constituents (butt & geuder, 2001); thus, neither annotators nor automatic systems can rely on syntactic criteria to identify lvcs and must use semantic criteria instead. the propbank guidelines for lvc annotation developed as part of this research therefore focus on the semantic nature of arguments (bonial & palmer, 2016).3 this helped to ground the guidelines in theoretical research on english lvcs, which generally assumes that the semantic content of the construction stems from the noun, while the verb provides the syntactic scaffolding for the noun to act as the main predicate (butt, 2003; grimshaw & mester, 1988). accordingly, the arguments, including the syntactic subject of the verb, will carry the semantic roles of the noun. thus, for lvc annotation, the annotators are instructed to compare the fit of the semantic roles listed for the verb’s roleset to that of the noun’s roleset. figure 2 gives an overview lvc designation heuristics. précis of take a look at this! form, function, and productivity of english lvcs 7 consider the earlier sentence from example (6), involving make_offer. annotators would examine the rolesets to decide which is more fitting: make-01 arg0: creator arg1: creation arg2: created-from selecting the above make-01 roleset requires designation of the following roles: (9) [ne electric]arg0 made [an offer of $2 billion]arg1 [to acquire ps]arg-??? offer-01 arg0: entity offering arg1: commodity arg2: price figure 2: flow chart for determining lvc status. colorado research in linguistics, volume 25 (2021) 8 while selecting the above offer-01 roleset requires designation of the roles shown below: (10) [ne electric]arg0 made an offer of [$2 billion]arg2 [to acquire ps]arg1 the verb-centric annotation in 9 does not capture the semantics of the offer event, instead focusing on make as if it were a heavy verb. thus, an argument lacks an appropriate tag in the roleset (indicated by ‘???’). the noun roleset offer-01 provides the appropriate roles to label each argument and allows for the arg2 price to be captured, as seen in annotation 10. these considerations are evidence that the semantic roles stem from the noun relation and the usage should be annotated as an lvc. the success of these guidelines has been demonstrated in the high agreement rates between annotators.4 on a task composed solely of the most likely lvs (give, have, take, make, do), the agreement rate between two seasoned annotators was 93.8%.5 as proof of the quality of propbank’s annotations, an automatic lvc detection system has been trained on the ontonotes 4.99 corpus (weischedel et al., 2011), of which propbank is one layer. this system achieves an f-score of 89% (chen, bonial & palmer, 2015) when tested on the same data used by tu and roth (2011) to establish the previous state-of-the-art system with an f-score of 86.3%.6 when tested on the more realistic and challenging ontonotes corpus containing 1,768 lvc instances, the system achieves an f-score of 80.7% (see system implementation and evaluation details in (chen, bonial & palmer, 2015)). 5. corpus & acceptability studies 5.1 corpus study: why do lvcs exist alongside semantically similar lexical verbs? in addition to providing valuable training data, the propbank corpus provides crucial information for investigating why english lvcs exist in the language alongside semantically similar counterpart lexical verbs. this question is important for an understanding of language use more generally, since the maintenance of both forms runs contrary to assumptions that two competing forms are rarely maintained unless they serve distinct purposes. the propbank corpus was used to analyze approximately 2,000 lvc annotations and 10,000 counterpart synthetic verb annotations. the aim was to discover evidence of what contexts call for the use of an lvc over a lexical verb. précis of take a look at this! form, function, and productivity of english lvcs 9 corpus study summary propbank annotations for those lvcs with counterpart lexical verbs (e.g., give a presentation), including all semantic roles and modifier arguments (e.g., temporal, locative, manner), were tabulated and compared to tabulations of roles and modifier arguments for counterpart lexical verbs (e.g., present). corpus study results the corpus study shows that lvcs are associated with significantly more modifiers (m = 1.17, sd = 0.45) than verb counterparts ((m = 0.60, sd = 0.23), t(36) = 4.93, p < .001). data were checked for normality and homogeneity of variance, and met the required statistical assumptions. due to the diverse and often idiomatic usage of these words, they were deemed to be independent, rather than paired, samples. however, using a paired-samples comparison yielded the same results, with lvcs having significantly more modifiers than their specific verb counterparts (t(18) = 5.60, p < .001). chart 1 shows the average number of modifiers across the two predicate types. verb counterpartslvcs predicate type 1.4 1.2 1.0 0.8 0.6 0.4 0.2 0.0 a ve ra ge n um be r of m od ifi er s chart 1: number of modifiers (mean ± 95% ci) associated with lvcs and verb counterparts. *** = p < 0.001. *** colorado research in linguistics, volume 25 (2021) 10 there is also greater variety to the types of modifiers seen with lvcs than counterpart synthetic verbs: nouns are compatible with descriptive elements that can only be expressed periphrastically with synthetic verbs. this can be observed in the following usage example of an lvc with evaluative modification, followed by an invented example that attempts a reformulation using the corresponding verb: (11) we had a really good laugh… (12) ?we laughed really well… example 12 shows that the transformation of an evaluative nominal modifier into an adverbial modifier changes the meaning into what will likely be interpreted as a somewhat awkward and ambiguous manner modifier. in general, speakers can exploit lvcs to express detailed descriptions of an event using nominal modification, which is more flexible than verbal modification. thus, this corpus study provided distributional evidence that the ease and variety with which lvcs can be modified, in order to provide nuanced and detailed descriptions of events, is the primary motivating factor for their use. since the lexical verb is not necessarily a viable alternative in the denotation of certain detailed events, the two forms are not competing in these cases, and therefore both forms are maintained. this study also reveals a characteristic of lvcs important for their automatic detection: they consistently include some modification of the noun. this characteristic is overlooked in nlp, as many lvc detection approaches operate under the assumption that only a determiner will occur between the verb and noun (e.g., stevenson et al., 2004). 5.2 acceptability study: what are the constraints on lvc productivity? although the propbank corpus was instrumental in providing training data for the detection of most lvcs, infrequent and novel lvcs unseen in the training data remain a challenge. one could assume that the same lv can combine with any member of a family of semantically related nouns, using information on semantic similarity from wordnet (fellbaum, 1998) or framenet (fillmore et al., 2002) to establish these families. however, this approach would over-generate possibilities: lvcs are not fully productive, given that some combinations are simply unacceptable. thus, this research aims to establish why certain combinations are acceptable while others are not, and to précis of take a look at this! form, function, and productivity of english lvcs 11 frame this knowledge in a way that can be usefully injected into nlp systems reliant on statistical patterns. emergent grammar predicts that novel constructions are extended by analogy to highfrequency, existing constructions, and this has been demonstrated by bybee and eddington (2006) in their research on spanish becoming constructions. the authors conducted acceptability studies of somewhat synonymous verbs, roughly meaning become, combined with different adjectives. they examined the acceptability of both attested and unattested combinations. it was found that speakers judged low-frequency or unattested combinations to be acceptable only if they were semantically very similar to an existing, high-frequency combination. intuitively, it seems plausible that other semi-productive constructions, like lvcs, would also be extended in this fashion. thus, the hypothesis was adapted to examine its validity with respect to english lvcs: frequency hypothesis: speakers will find novel or very low-frequency lvcs acceptable if they are semantically similar to an attested, highly frequent lvc. in accordance with the findings of bybee & eddington (2006), the expectation is that speakers will find very low-frequency lvcs that are similar to a strongly entrenched, high-frequency exemplar to be more acceptable than those that are similar to a weaker, low-frequency exemplar. to test this hypothesis, it was necessary to establish what families of semantically similar lvcs exist, and what the frequency is of each unique lvc within a family. although propbank lvc annotations were used to determine common lvs of focus, the corpus was not large enough to gain a full picture of lvc usage in english. thus, occurrences of the frequent lvs give, have, make and take were collected from the gigaword corpus of over 1.7 billion words (graff & cieri, 2003). framenet membership was used to determine what lvcs were semantically similar. after computational analysis of frequencies and some manual filtering of false positives, a fairly comprehensive picture of families of lvcs emerged, including very low-frequency lvcs to be tested for acceptability that had either a low-frequency exemplar within their family, or a highfrequency exemplar. “very low-frequency” lvcs occur 20-50 times in gigaword, “lowfrequency” lvcs occur 100-200 times, “high-frequency lvcs occur 2000-3000 times. these frequency bands were decided upon after careful analysis of histograms displaying how many colorado research in linguistics, volume 25 (2021) 12 unique collocations occur at which frequencies across gigaword. examples of the two types of lvcs compared for acceptability here—very low-frequency lvcs with only a low-frequency exemplar in the semantically similar family, and very low-frequency lvcs with only a highfrequency exemplar in the semantically similar family—are shown in tables 1 and 2. lvc frequency make realization 5 make inference very low-frequency test lvc 20 make deduction 24 make guess low-frequency exemplar lvc 102 lvc frequency have dread very low-frequency test lvc 24 have terror 33 have apprehension 107 have fear high-frequency exemplar lvc 2342 the very low-frequency test lvcs were presented to 125 participants on amazon’s mechanical turk in the context of a sentence exemplifying that lvc taken from gigaword. participants provided a rating for the rare lvc by clicking in a box with just two endpoints marked as “odd,” on one end, or “perfectly fine.” the survey interface is shown in figure 3. note that participants are never presented with low or high-frequency exemplar lvcs, they are solely judging the table 1: a low-frequency family of lvcs detected in gigaword: most tokens fall into the low-frequency band of 100-200 instances. the nouns of the lvcs in these families share a framenet frame. table 2: a high-frequency family of lvcs detected in gigaword: most tokens fall into the high-frequency band of 2,000-3,000 instances. précis of take a look at this! form, function, and productivity of english lvcs 13 acceptability of the very low-frequency test lvcs, so they are in no way primed to consider similarity to any existing, higher-frequency lvc. acceptability study summary native english speakers judge the acceptability of very low-frequency lvcs that are either semantically similar to a low-frequency or high-frequency exemplar, and the levels of acceptability will be compared across the two types. acceptability study results it was found that the frequency of the exemplar lvc does significantly correlate with the acceptability of the very low-frequency test lvc, but in a manner opposite of what was predicted: very low-frequency lvcs that are semantically similar to low-frequency lvc exemplars are significantly more acceptable than very low-frequency lvcs that are semantically similar to highfrequency exemplars (χ2(1)=5.64, p=0.01751).7 the most plausible explanations for this surprising result were explored: statistical preemption and semantic bleaching. statistical preemption is a cognitive process in which speakers implicitly infer from consistently hearing a formulation, b, in a context where one might have heard a semantically figure 3: question interface viewed by participants. colorado research in linguistics, volume 25 (2021) 14 related alternative formulation, a, that b is the appropriate formulation and a is not appropriate (e.g., suttle & goldberg, 2011). here, the highly frequent exemplar lvc blocks the use of the semantically similar, very low-frequency lvc. thus, when speakers judge the acceptability of a relatively unfamiliar, very low-frequency lvc (e.g., give a stare) that is clearly quite similar to a very familiar, high-frequency construction (e.g., give a look), the unfamiliar lvc seems odd, given that the familiar construction could have been used. on the other hand, when speakers judge the acceptability of a very low-frequency lvc (e.g., take a departure) that is similar to a lowfrequency construction (e.g., take an exit), the low-frequency exemplar is likely familiar enough that it can serve as the basis of analogical extension of the family to the new, unfamiliar member, but the exemplar isn’t so frequent that it blocks the unfamiliar alternative. semantic bleaching is a process in which a high-frequency phrase or construction has become entrenched to the point where speakers no longer analyze the individual components of the expression, instead treating it as an unanalyzed whole (hopper & traugott, 1993). when speakers judge the acceptability of an unfamiliar lvc, they are comparing it to previously experienced constructions for similarity, in order to decide if the unfamiliar construction should be added to an existing family of constructions. however, the relationship of semantic similarity between the unfamiliar construction (e.g, have dread) and the familiar construction (e.g., have fear) is not clear after semantic bleaching has taken place because speakers no longer analyze the individual components of the high-frequency expression, precluding recognition of the semantic similarity between those components (e.g., dread and fear). therefore, the unfamiliar construction is not accepted as a member of the existing family. the first interpretation involving statistical preemption seems more likely, given trends in this data and a survey of related work on both statistical preemption and semantic bleaching. lending further evidence for this interpretation, it was found that very low-frequency lvcs that were from semantically similar families with more than one high-frequency member were significantly less acceptable than those that came from a family with only one higher-frequency member (χ2(1)=4.84, p=0.02775). essentially, in families with more than one high-frequency member, there are two entrenched formulations that can block the less frequent alternative formulation. chart 2 exemplifies such a family. précis of take a look at this! form, function, and productivity of english lvcs 15 furthermore, while the process of statistical preemption has been demonstrated quite convincingly in experimental settings (e.g., boyd, ackerman & kutas, 2012), the process of semantic bleaching has not. in fact, although the process of semantic bleaching seems very plausible from a diachronic perspective, a growing body of research in psycholinguistics and cognitive science has demonstrated that speakers do recognize and analyze the individual lexical items within lvcs and idiomatic expressions (e.g., cutting & bock, 1997; wittenberg & piñango, 2011), contrary to semantic bleaching predictions. thus, this research demonstrates that even relatively low-frequency constructions can serve as the basis of analogical extension, which evidences the emergent grammar assumption that speakers retain details of each token of linguistic experience. furthermore, there is a certain point where the higher frequency of a semantically similar form may block another variant, rather than encouraging analogical extension, thereby providing evidence for the somewhat overlooked role of statistical preemption in the process of extending constructions. to take advantage of these findings in nlp, we can use the frequency signatures of lvc families to predict which families will block or allow new members, effectively modeling patterns of lvc productivity and acceptability. chart 2: family with two high-frequency lvcs, in which both give permission and give approval block the use of the rare variant give sanction. colorado research in linguistics, volume 25 (2021) 16 6. conclusions, interdisciplinary contributions this work draws upon research strands from linguistics, cognitive science, and computer science to make several theoretical and practical contributions. first, the development of a valuable resource of annotated english lvcs allows for continued linguistic research on these constructions and provides training data for the automatic detection of lvcs.8 next, this research provides quantitative evidence demonstrating why two seemingly competing forms, lvcs and counterpart lexical verbs, are maintained in the language. finally, this work provides evidence for the role of statistical preemption in extending semi-productive constructions. this informs our picture of how grammar is constructed as a whole, and provides a computational framework for estimating the likelihoods of productivity and acceptability. this information will be instrumental in the automatic detection of novel lvcs, unseen in training data. as a whole, this research has demonstrated the promising ways in which the various disciplines of (psycho)linguistics, cognitive science and computer science can be brought together to identify a linguistic phenomenon, understand its function in the language, and begin to understand how humans can, and perhaps computers should, process and extend the phenomenon. references bates d, m. maechler, b. bolker and s. walker. 2014. lme4: linear mixed-effects models using eigen and s4. r package version 1.1-7, http://cran.r-project.org/package=lme4. bonial, claire & martha palmer. 2016. comprehensive and consistent propbank light verb annotation. in proceedings of the tenth international conference on language resources and evaluation (lrec'16), pp. 3980-3985. bonial, claire, julia bonn, kathryn conger, jena d. hwang & martha palmer. 2014. propbank: semantics of new predicate types. in proceedings of the ninth international conference on language resources and evaluation (lrec’14), pp. 3013-3019. boyd, jeremy k., farrell ackerman and marta kutas. 2012. adult learners use both entrenchment and preemption to infer grammatical constraints. in proceedings of the 2012 ieee international conference on development and learning and epigenetic robotics (icdl). butt, miriam. 2003. the light verb jungle. in g. aygen, c. bowern & c. quinn (eds.) papers from the gsas/dudley house workshop on light verbs. cambridge, harvard working papers in linguistics, p. 1-50. précis of take a look at this! form, function, and productivity of english lvcs 17 butt, miriam, and wilhelm geuder. 2001. on the (semi)lexical status of light verbs. in semilexical categories: on the content of function words and the function of content words, ed. norbert corver and henk van riemsdijk. 323-370. berlin: mouton de gruyter. bybee, joan. 2006. from usage to grammar: the mind's response to repetition. language 82(4): 711-733. bybee, joan. 2010. language, usage and cognition. cambridge: cambridge university press. bybee, joan, and d. eddington. 2006. a usage-based approach to spanish verbs of ‘becoming’. language, 82(2), 323–355. cattel, ray. 1984. syntax and semantics: composite predicates in english. orlando: academic press, inc. chen, wei-te, claire bonial & martha palmer. 2015. english light verb construction identification using lexical knowledge. in proceedings of twenty-ninth conference on artificial intelligence (aaai-15). cutting, j. cooper and kathryn bock. 1997. that’s the way the cookie bounces: syntactic and semantic components of experimentally elicited idiom blends. memory & cognition, 25(1): 57-71. davies, m. 2008-. the corpus of contemporary american english (coca): 425 million words, 1990-present. available online at http://www.americancorpus.org. fellbaum, c. (ed.) 1998. wordnet: an electronic lexical database. mit press. fillmore, charles j., christopher r. johnson, and miriam r.l. petruck. 2002. background to framenet. international journal of lexicography, 16(3):235-250. goldberg, adele. 1995. constructions: a construction grammar approach to argument structure. chicago: university of chicago press. goldberg, adele. 2006. constructions at work: the nature of generalization in language. new york: oxford university press. goldberg, adele. 2011. corpus evidence of the viability of statistical preemption. cognitive linguistics 22 1: 131-154. graff, david and christopher cieri. 2003. english gigaword. linguistic data consortium, philadelphia. grice, h.p. 1975. logic and conversation. in p. cole and j. morgan (eds.) syntax and semantics, vol.3. academic press. reprinted as ch.2 of grice 1989, 22–40. colorado research in linguistics, volume 25 (2021) 18 grimshaw, j., and a. mester. 1988. light verbs and θ-marking. linguistic inquiry, 19(2):205– 232. hopper, paul j. 1998. emergent grammar. in the new psychology of language: cognitive and functional approaches to language structure, michael tomasello (ed.), 155-75. mahwah, nj/london: lawrence erlbaum associates. hopper, paul j. and elizabeth c. traugott. 1993. grammaticalization (=cambridge textbooks in linguistics.) cambridge: cambridge university press. hwang, jena d, archna bhatia, claire bonial, aous mansouri, ashwini vaidya, nianwen xue, and martha palmer. 2010. propbank annotation of multilingual light verb constructions. in proceedings of the linguistic annotation workshop held in conjunction with acl-2010. uppsala, sweden, july 15-16, 2010. jespersen, otto. 1942. a modern english grammar on historical principles. part vi: morphology. with assistance of p. christopherson, n. haislund, & k. schibsbye. london: goerge allen & unwin. copenhagen: ejnar munksgaard. nickel, gerhard. 1978. complex verbal structures in english. studies in descriptive english grammar, 63-83. palmer, martha, daniel gildea, and paul kingsbury. 2005. the proposition bank: an annotated corpus of semantic roles. computational linguistics, 31(1):71– 106. passonneau, rebecca. 2004. computing reliability for coreference annotation. piñango, maria m, jennifer mack and ray jackendoff. 2006. semantic combinatorial processes in argument structure: evidence from light-verbs. in proceedings of the thirty-second annual meeting of the berkeley linguistics society. postal, paul.m. 1974. on raising: one rule of english grammar and its theoretical implications (vol. 5). the mit press. r development core team. 2011. r: a language and environment for statistical computing. r foundation for statistical computing, vienna, austria. isbn 3-900051-07-0, url http://www.r-project.org/. suttle, laura and adele e. goldberg. 2011. partial productivity of constructions as induction linguistics 49 6: 1237-1269. précis of take a look at this! form, function, and productivity of english lvcs 19 tu, yuancheng, and dan roth. 2011. learning english light verb constructions: contextual or statistical. in proceedings of the worshop on multiword expressions: from parsing and generation to the real world (mwe 2011), pp. 31-39. weischedel, r., e. hovy, m. palmer, m. marcus, r. belvin, s. pradhan, l. ramshaw and n. xue. 2011. ontonotes: a large training corpus for enhanced processing. in olive, j.; christianson, c.; and mccary, j., eds., handbook of natural language processing and machine translation. springer. wierzbicka, anna. 1982. why can you have a drink when you can’t *have an eat? language, 58(4):753–799. wittenberg, eva & maria m. piñango. 2011. processing light verb constructions. the mental lexicon 6:3, 393–413. zarco, m. a. 1999. interlingual representation of complex predicates in a multilingual approach: the problem of lexical selection. saint-dizier, p. (ed.) predicative forms in natural language and in lexical knowledge bases. dordecht: kluwer academic publishers. zipf, george k. 1949. human behavior and the principle of least effort. addison-wesley. colorado research in linguistics, volume 25 (2021) 20 endnotes 1 unacceptable usages are preceded by ‘*’; questionable usages by ‘?’. 2 http://verbs.colorado.edu/propbank/framesets-english-aliases 3 http://verbs.colorado.edu/propbank/epb-annotation-guidelines.pdf 4 a common measure of the quality and reliability of an annotation schema (e.g., passoneau, 2004). 5 http://verbs.colorado.edu/propbank/ita/webtext-p25-sel-lightverb.html 6 f-score is the harmonic mean of a classification system’s precision and recall. 7 r (r core team, 2012) and lme4 (bates, maechler & bolker, 2012) were used to perform a linear mixed effects analysis of the relationship between the subject’s naturalness rating of the very lowfrequency lvcs and the frequency band of the exemplar lvc. 8 the lvc annotations have been released as part of the “bolt english propbank and sense - discussion forum, sms/chat, and conversational telephone speech” annotated corpus, ldc2020t21 (https://catalog.ldc.upenn.edu/ldc2020t21).