#1424. CELEX2 ([2013-03-21-1]) ǾҲ𤷤ǡ١DzƤߤ褦ȹͤVersion 2 ǿ˲ä줿 (English Frequency, Syllables) Υ֥ǡ١ˤꡤѸǺǤ¿פΥ
ϡCELEX2 ΤȤˤʤäƤ륳ѥΤΤ7.26%130äե֥ѥФ줿٤Ǥꡤ٤ǤϤʤȡ٤ˤΤǤ롥Ĥޤꡤäդˤ뤢ñ٤⤱Сʬñ˴ޤޤ벻פ٤⤯ʤȤȤǤ롥㤨Сof "Ov" (= /ɒv/) ɽ벻ϡ4̤٤Ǥ롥ʤ̵ͭϹθ٤Ƥ롥
ʲΥꥹȤ˵벻ɽϡIPA ǤϤʤ CELEX ͤäɽʤΤǡбɽƤ

Ǥϡʲ˥ɽǥȥå50̤ޤǤǺܤ롥٤ñβפΤޤ̤ȿǤƤơޤꤪ⤷ɽǤϤʤΩĤȤ⤢뤫⤷ʤ
| Rank | Syllable | Frequency |
|---|---|---|
| 1 | eI | 72971 |
| 2 | Di: | 60967 |
| 3 | tu: | 31446 |
| 4 | Ov | 30108 |
| 5 | In | 29906 |
| 6 | &nd | 28709 |
| 7 | aI | 23822 |
| 8 | lI | 19728 |
| 9 | @ | 19566 |
| 10 | rI | 14356 |
| 11 | ju: | 12598 |
| 12 | dI | 12465 |
| 13 | D&t | 12118 |
| 14 | It | 11504 |
| 15 | wOz | 10834 |
| 16 | fO:r* | 9778 |
| 17 | Iz | 9517 |
| 18 | tI | 9161 |
| 19 | fO | 9042 |
| 20 | Sn, | 8969 |
| 21 | hi: | 8928 |
| 22 | r@n | 8638 |
| 23 | bi: | 8505 |
| 24 | bI | 7936 |
| 25 | nI | 7068 |
| 26 | wID | 7046 |
| 27 | On | 7030 |
| 28 | &z | 6919 |
| 29 | O:l | 6569 |
| 30 | h&d | 6240 |
| 31 | E | 6165 |
| 32 | bl, | 6021 |
| 33 | sI | 5836 |
| 34 | @U | 5824 |
| 35 | t@r* | 5687 |
| 36 | &t | 5652 |
| 37 | hIz | 5564 |
| 38 | bVt | 5416 |
| 39 | mI | 5397 |
| 40 | s@ | 5391 |
| 41 | nOt | 5357 |
| 42 | D@r* | 5339 |
| 43 | I | 5283 |
| 44 | tId | 5259 |
| 45 | DeI | 5162 |
| 46 | IN | 5063 |
| 47 | t@ | 5053 |
| 48 | s@U | 4974 |
| 49 | baI | 4894 |
| 50 | h&v | 4769 |
Please note that the English corpus used by CELEX for deriving these frequencies contains only 7.3% spoken material. This means there is a rather tenuous relationship between the full frequency figures, which are based on written forms, and the syllable frequencies, which merely refer to phonemic conversions of these graphemic transcriptions. Of course it could be argued that frequencies of syllables, as lexical sub-units, are less liable to get skewed from differences in medium than full words, but it has to be taken into account that NO FIRM EVIDENCE ABOUT SPOKEN FREQUENCIES can be derived from these data.
#13. ѹΥѥ֤ ye äƤ椷 ([2009-05-11-2]) ǡye ye 괧 the Ѥ뵼ŪֻˤĤƿ줿
þ (thorn) y Ȥλˤ뺮ƱѸ鸫줿κ𤬤дΤ þ षƤǤ롥þ ѤƤäΤϡ#1329. Ѹˤˤ eth, thorn, <th> ([2012-12-16-1]) #1330. Ѹˤ eth, thorn, <th> ([2012-12-17-1]) dzǧ褦ˡHelsinki Corpus λʬˤME4 (1420--1500) ʹߤǤ롥˸ƱơŪ괧 ye ϶ѸäƤ٤ƤOED Ȥȡye λѤѸ줫17ˤơȤ롥
ǤϡѸ줫ѸˤơŪˤɤ ye Ѥ줿ΤĴ٤뤿 PPCME2, PPCEME, PPCMBE POSե뷲 "ye/D" ƤߤME1ΤߡEModE1259㡤LModE5㤬äƥѥϤ褽130졤180졤100줫ʤ뤬ͤȤ⡤ȤƤ롥Ѹǵ˸쵤˿ȤȤǤ롥PPCEME 1259Τ975ϡThe Journal of George Fox (1673--74) Ȥ1ʤǤ롥ۤˤ10ʾ帽ƥȤ4ĤΤߤǡĤ20ƥȤ˾㤺ĻФäƤˤʤȤʬۤǤϤ롥δˤȤϡ̣ήԤȤä
ɥˬ줿ݤˡ145 Fleet St Ϸޥѥ "Ye Olde Cheshire Cheese" 42 Ludgate Hill "Ye Olde London" δĤƤƤǰʤ餳ǥդ뵡Ϥʤäɤ⡤̤Υѥ֤ǤϰաʤǤϤʤˤޤ

ñ٤˴ϢBetty Phillips ʤɡˤǡCELEX Ȥåǡ١ѤƤΤ뤳Ȥ롥꤫äƤ븦ǡ祳ѥ˴ŤǤפɬפˤʤäΤǡߤ350ɥ뤹뤳ιʥǡ١ꤷƤߤǤ2ǤǤꡤCELEX2 ȤƹǤ롥ʤʤͽۤƤʤäꤷ CD-ROM ˤϡLDC99T42 Ȥǡ١ޤޤƤˤ tagged Brown Corpus, Wall Street Journal, Switchboard tagged ʤ Treebank ϤΥѥäƤ롥
ơCELEX2 ˤϡѸä˴ؤʣΥǡ١ǼƤ롥줾Υǡ١ˤϡˡᡤ֡γƴ顤Ф (lemma) 뤤ϸ (wordform) ȤˡѥǤξǼƤ롥Ūˤϡ11Υǡ١ѲǽǤ롥
ect (English Corpus Types)
efl (English Frequency, Lemmas)
efs (English Frequency, Syllables)
efw (English Frequency, Wordforms)
eml (English Morphology, Lemmas)
emw (English Morphology, Wordforms)
eol (English Orthography, Lemmas)
eow (English Orthography, Wordforms)
epl (English Phonology, Lemmas)
epw (English Phonology, Wordforms)
esl (English Syntax, Lemmas)
Ф줢뤤ϸȤ token ٤μФ˶ǡ١ȤǧǹºݤˤϡޤޤƤμ϶äۤ˭٤ǡ11Υǡ١٤Ƥ碌եɿϤΤ250ʾ˵ڤ֡Կ efl 52,447ԡefw 160,595ԤȤ礵Ѥ SQLite DB 館顤̤ˤ90MBĶƤޤä
CELEX2 ΥϡˤĤƤ Oxford Advanced Learner's Dictionary (1974) ڤ Longman Dictionary of Contemporary English (1978) ǤꡤپˤĤƤ 1790줫ʤ COBUILD/Birmingham corpus Ǥ롥Υѥιϡ1660 (92.74%) եѥ130 (7.26%) äեѥǡԤ284ƥȤΤ44ƥ (15.49%) ꥫѸǤ롥ΥꥫѸϤۤȤɤꥹѸֻľƤ뤳Ȥդ
CELEX2 ˤ "lemma" ϡʲ5˰¸롥
(1) orthography of the wordforms: peek vs peak
(2) syntactic class: meet (adj.) vs meet (adv.)
(3) inflectional paradigm: water (v.) vs water (n.)
(4) morphological structure: rubber (someone or something that rubs) vs rubber (the elastic substance)
(5) pronunciation of the wordforms: recount [ˈriː-kaʊnt] vs recount [rɪ-ˈkaʊnt]
äơ̾ۤʤ lexeme Ȥư bank (ڼˤ bank ʶԡˤʤɤϡCELEX2 ǤƱ lemma ȤưƤΤդɬפǤ롥
Τ褦 CELEX2 ˶Ϥʸ٥ǡ١¾ˤٸ˻ǡ١ġ¸ߤ롥ܥ֥ǿ줿ΤȤƤϡfrequency statistics lexicology γƵ䡤ä˰ʲεͤˤʤ
#308. Ѹκѱñꥹȡ ([2010-03-01-1])
#607. Google Books Ngram Viewer ([2010-12-25-1])
#708. Frequency Sorter CGI ([2011-04-05-1])
#1159. MRC Psycholinguistic Database Search ([2012-06-29-1])
Baayen R. H., R. Piepenbrock and L. Gulikers. CELEX2. CD-ROM. Philadelphia: Linguistic Data Consortium, 1996.
#1413. Ѹ3ʣ -s ([2013-03-10-1]) ε³εǤϡPPCEME ˤ븡ǡ3ʣ -s 50ۤɼФȤǤȽҤ٤ʸ̮ʤȤȤ52㤬ǧ줿ʥǡΥƥȥեˡ
PPCEME ǤϡE1 (1500--1569), E2 (1570--1639), E3 (1640--1710) 3ʬƤ뤬ζʬȤ3ʣ -s ȰʲΤ褦ˤʤʳƴΥѥ⼨ˡ
| Period | Tokens | Wordcount |
|---|---|---|
| E1 (1500--1569) | 13 | 567,795 |
| E2 (1570--1639) | 18 | 628,463 |
| E3 (1640--1710) | 21 | 541,595 |
| Total | 52 | 1,737,853 |
and after them comys mo harolds,
Here comes our Gossips now,
Now in goes the long Fingers that are wash't Some thrice a day in Vrin,
ơLass (166) 3ʣ -s ˤĤƴϢڤĤΤǡҲ𤷤ƤLass ϡ3ʣ -s εˤĤơñ٤л٤줿ΤΡŤȹͤƤ褦
The {-s} plural appears considerably later than the {-s} singular, and if it too is northern (as seems likely), it represents a later diffusion. The earliest example cited by Wyld ([History of Modern Colloquial English] 346) is from the State Papers of Henry VIII (1515): 'the noble folk of the land shotes at hym'. It is common throughout the sixteenth and seventeenth centuries as a minority alternant of zero, and persists sporadically into the eighteenth century.
1617̤ƹԤʤƤȤȤϡ嵭 PPCEME dzΤǧ줿ʤѸС PPCMBE 18ʹߤξĴ٤Ƥߤȡ6äΤοȴǰξǾȤפɤΤޤޤƤ[2012-06-14-1]ε#1144. ѸˤפפȡˡѸǤ3ʣ -s ϳ̵˶ᤤȹͤƤ褵
Lass, Roger. "Phonology and Morphology." 1476--1776. Vol. 3 of The Cambridge History of the English Language. Ed. Roger Lass. Cambridge: CUP, 1999. 56--186.
Ѹʹߤ apostrophe s ϡ̾ˤĤȤߤʤϡ̾ˤĤܸ (enclitic) ȤߤʤۤΤǤ롥ȤΤϡthe king of England's daughter Τ褦° (group genitive) ȤƤˡǧ뤫Ǥ롥apostrophe s ϡ³˷Ūñ̤Ȥϡ췲³Ūñ̡ʤܸ (clitic) Ȥߤʤɬפ롥
apostrophe s εȹͤ -es ϡѸˤϡΤ̾ˤĤä줬̾ܤ°ʤˡΤϤʤηϲäΤѸˤǤ˵ƤǤ롥
Ĥθ (Janda) ˤС°ʤؤȯŸʳǡ#819. his °ʡ ([2011-07-25-1]) ȤƺѤΤǤϤʤȤñ㲽ƼС(1) °ʸ -es ȿ;̾ñ° his Ȥ̵ƱȤʤ¤ȡ(2) ľ̾̾ȤƤ his °ʤȤ2ؤäơΤ褦㼰ǽȤʤäΤǤϤʤȤ
king his doughter : king of England his doughter = kinges doughter : X
X = king of Englandes doughter
Allen ϤƱդʤPPCME2 䤽¾ѸƥȤͿ뤢Ƥ̡his °ʤȤʤäƷ°ʤȤ븫ˤϡڵ塤̵Ȥ롥Allen ϡȤ櫓"attached genitive" Ū -es °ʡˤ "separated genitive" his °ʡˤȤδ֤ˡĶ˱Ƥʬ۾κʤȤˡѸ his °ʤ "just an orthographical variant of the inflection" (118) Ǥȷ롥
Ǥϡ°ʤȯã his °ʤȤΤǤϤʤäȤȡ¾ˤɤΤ褦ʷꤨΤAllen ϤȤơ"the gradual extension of the ending -es to all classes of nouns, making what used to be an inflection indistinguishable from a clitic" (120) ƤƤ롥14ޤǤ°ʸΧ -(e)s 褦ˤʤꡤ줬ϤȤƤǤϤʤ̵Ѳܸª˻äΤǤϤʤȤޤǽη°ʤϡThe grete god of Loves name (Chaucer, HF 1489) þe kyng of Frances men (Trevisa's Polychronicon, VIII, 349.380) ˸褦ʡм of ȼä귿Ǥꡤʣ̾ȤǤ褦ɽǤ롥줬Ȳ졤ľ˽ͭ -(e)s ĤȤΤϤޤäԻĤǤϤʤ
Allen ϡ16Ⱦ鸽 his °ʤʿŪ her °ʤ their °ʤˤĤƤϡǤ-(e)s ˤ뷲°ʤΩʬ (metanalysis) η̤ǤꡤŪɽˤʤȸƤ롥ΰʬϤ"spelling pronunciation" ʤ "spelling syntax" (124) ȸڤƤΤ̣
Allen ηѤ褦 (124)
A closer examination of the relationship between case-marking syncretism and the rise of the 'group genitive' than has previously been carried out provides evidence that the increase in syncretism led to the reanalysis of -es as a clitic. There is evidence that this change of status from inflection to clitic was not accomplished all at once; inflectional genitives coexisted with the clitic genitive in late ME and the clitic seems to have attached to conjoined nouns and appositives before it attached to NPs which did not end in a possessor noun. The evidence strongly suggests that the separated genitive of ME did not serve as the model for the introduction of the group genitive, and I have suggested that the separated genitive was an orthographic variant of the inflectional genitive, but that after the group genitive was firmly established there were attempts to treat it as a genitive pronoun.
Allen Appendix I ˤơMustanoja (160) his °ʤȤƵƤűѸѸ줫[2011-07-25-1]ǵˤ¿路ǤƤ롥
ʤ°ʤˤĤƤϡBaugh and Cable (241) ˤñʸڤ롥
Janda , Richard. "On the Decline of Declensional Systems: The Overall Loss of OE Nominal Case Inflections and the ME Reanalysis of -es as his." Papers from the Fourth International Conference on Historical Linguistics. Ed. Elizabeth C. Traugott, Rebecca Labrum, and Susan Shepherd. Amsterdam: John Benjamins, 243--52.
Allen, Cynthia L. "The Origins of the 'Group Genitive' in English." Transactions of the Philological Society 95 (1997): 111--31.
Mustanoja, T. F. A Middle English Syntax. Helsinki: Société Néophilologique, 1960.
Baugh, Albert C. and Thomas Cable. A History of the English Language. 5th ed. London: Routledge, 2002.
ε#1414. shew show (1) ([2013-03-12-1]) ³ԡ Helsinki Corpus ѤƽѸޤǤ shew show ʬۤĴϸѸˤʬۤ PPCMBE (Penn Parsed Corpus of Modern British English; see [2010-03-03-1]) ˤäƴñĴ
PPCMBE ϡ1700ǯ1914ǯޤǤ948,895ΥѥǤ롥70ǯĤ3ʬФ첽줿 pos ե뷲оݤ˸뤳Ȥ shew show token 夲̤ϰʲ̤ꡥ
| shew | show | ||
|---|---|---|---|
| 1700--1769 | 80 | 25 | 298,764 |
| 1770--1839 | 79 | 86 | 368,804 |
| 1840--1914 | 17 | 162 | 281,327 |
This word is frequently written shew; but since it is always pronounced and often written show, which is favoured likewise by the Dutch schowen, I have adjusted the orthography to the pronunciation.
Ĥޤꡤspelling_pronunciation ʤ pronunciation_spelling ȤȤˤʤΤshow ۤɤιٸǤΤ褦ʰŪʲѤȤΤԻĤˤפ뤬Ѹ衤ȤϤ show ϹԤʤƤȤ¤طʤˤäȤϡΤ˸Ƥ
ư show ˤϸŤ֤ shew 롥ˡΧʸʤɤˤϸΤΡߤǤϰŪˤϤޤꤪܤˤʤshew 18ޤֻͥǤꡤ19ȾޤǸȤƳƤ20ȾǤܤˤ뤳ȤäOED "show, v." θ츻εҤȤ褦
The spelling shew, prevalent in the 18th cent. and not uncommon in the first half of the 19th cent., is now obsolete exc. in legal documents. It represents the obsolete pronunciation (indicated by rhymes like view, true down to c1700) normally descending from the Old English scéaw- with falling diphthong. The present pronunciation, to which the present spelling corresponds, represents an Old English (? dialectal) sceāw- with a rising diphthong.
ΰ֤ͳϡűѸ scéawian (to look) η֤ͳ褹롥촴2첻Ĺ첻ؤȳ경 (smoothing) ݤˡȤȲĴ2첻Ǥкǽ첻Ӥ ē Ȥʤꡤ徺Ĵ2첻ǤкǸ첻ӡ̤Ȥ ō Ȥʤäshew ϢʤԤηǤϡ§Ūʲȯãˤꡤ/ʃjuː/ ϤϤshow ϢʤԤηȯ /ʃoʊ/ ִ뤳ȤˤʤäȾޤǤֻȤƤ shew ͥǤʤ顤ȯȤƤ show ̲ƤȤȤˤʤ롥ʤsew /soʊ/ ⡤ʿŪȯãη̤Ǥ롥
OED 츻ǤϡޤǤ shew ͥäȤȤä show ˨Ѹ줫ǧ (cf. MED "sheuen (v.(1))") Helsinki Corpus ˤꡤѸ줫ѸޤǤ shew vs show ̻ŪʬۤѤƤߤ褦ʥǡե "shew" and "show" in Helsinki Corpus ȡ
| shew | show | |
|---|---|---|
| M1 | 12 | 0 |
| M2 | 36 | 2 |
| M3 | 185 | 0 |
| M4 | 207 | 7 |
| E1 | 198 | 13 |
| E2 | 113 | 15 |
| E3 | 71 | 4 |
ɸˤĤ Baugh and Cable (247) ˿Ƥꡤܤ椤3ñʤ3ʣˤ -s ϡѸǤʤѸǤϡ#790. Ѹˤưʬۡ ([2011-06-26-1]) βϿޤǼ褦ˡľˡ3;ʣǤ -es ܤäѸɸѼˤ3ʣ -s ȤΤԻĤǤ롥ȤΤϡλʸؤʸȿǤɸѼǤϡѸ East Midland -e(n) ṳ̈ȤƤΥͽۤ뤷ºݤʬۤȤưŪ3ʣ -s ϳΤ Shakespeare Ǥ롥
ˤĤơBaugh and Cable (247) ϼΤ褦˻ŦƤ롥
Their occurrence is also often attributed to the influence of the Northern dialect, but this explanation has been quite justly questioned, and it is suggested that they are due to analogy with the singular. While we are in some danger here of explaining ignotum per ignotius, we must admit that no better way of accounting for this peculiarity has been offered. And when we remember that a certain number of Southern plurals in -eth continued apparently in colloquial use, the alternation of -s with this -eth would be quite like the alternation of these endings in the singular. Only they were much less common. Plural forms in -s are occasionally found as late as the eighteenth century.
ǡ3ʣ -s αƶǤϤʤȤ Baugh and Cable θϡWyld, History of Modern Colloquial English, p. 340 θڤäƤ롥षλ3ñ -s -th θؤʣˤŪӲФ̤ȹͤƤ롥ʤGörlach (89) ϡ餭Τñ餭ΤᤫͤƤ롥
ͻˤäơϤȤ⤢졤Ѹˤ3ʣ -s ʤ -th ʤ꤬ŪˤɤΤ餤٤ǸΤǧƤɬפ롥ǡThe Penn-Helsinki Parsed Corpus of Early Modern English (PPCEME) ˤꤶäȸƤߤ180ȤϤΥѥ3ʣ -s 50ۤɡ3ʣ -th 60ۤɤäʷ̥ƥȥեϺåˡǯʸ̮ʤɤξܺ٤ʬϤϤƤʤŵŪƤ
and all your children prayes you for your daly blessing.
but the carving and battlements and towers looks well;
then go to the pot where your earning bagges hangs,
as our ioyes growes, We must remember still from whence it flowes,
Ther growes smale Raysons that we call reysons of Corans,
now here followeth the three Tables,
And yf there be no God, from whence cometh good thynges?
First I wold shewe that the instruccyons of this holy gospell perteyneth to the vniuersal chirche of chryst.
and so the armes goith a sundre to the by crekes.
And to this agreith the wordes of the Prophetes, as it is written.
Also high browes and thicke betokeneth hardnes:
Baugh, Albert C. and Thomas Cable. A History of the English Language. 5th ed. London: Routledge, 2002.
Görlach, Manfred. Introduction to Early Modern English. Cambridge: CUP, 1991.
#1389. between θ츻 ([2013-02-14-1])#1393. between Ū۷֤˭٤([2013-02-18-1])#1394. between ΰ۷֤ʬ̻ۤŪѲ ([2013-02-19-1]) ³ơ LAEME Ѥ̻ŪѲʬۤĴ̤𤹤롥
Helsinki Corpus ˤ̻ŪĴ ([2013-02-19-1]) ξƱͤˡ¿ΰ۷֤ޤȤäơʳˤ첻ΰ㤤̵뤷2ʹߤλҲʤȡ⤷и첻ˤμȤ߹碌ܤlexel "between" ꤷƼФȤˡ241ĤΥȡȾȡ̤ʶʬ[2012-10-10-1]ε#1262. The LAEME Corpus ɽ (1)פǺѤΤƱˡǡȡʲǽǯ̡̤ν̤Ǥ롥
| PERIOD | nn | n | ne | x | xe | xn | xte | hn | he | tn | tx | txn | txe | ths | s | e | yn | zn | Sum |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| C12b | 18 | 1 | 2 | 7 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 28 |
| C13a | 23 | 4 | 19 | 6 | 4 | 4 | 0 | 9 | 14 | 0 | 1 | 0 | 1 | 0 | 0 | 0 | 0 | 0 | 85 |
| C13b | 20 | 3 | 23 | 2 | 1 | 3 | 4 | 1 | 0 | 0 | 0 | 1 | 0 | 2 | 1 | 1 | 1 | 1 | 64 |
| C14a | 5 | 13 | 28 | 9 | 2 | 2 | 0 | 0 | 0 | 3 | 1 | 0 | 0 | 0 | 0 | 1 | 0 | 0 | 64 |
| Sum | 66 | 21 | 72 | 24 | 7 | 9 | 4 | 10 | 14 | 3 | 2 | 1 | 1 | 2 | 1 | 2 | 1 | 1 | 241 |
| DIALECT | nn | n | ne | x | xe | xn | xte | hn | he | tn | tx | txn | txe | ths | s | e | yn | zn | Sum |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| N | 0 | 0 | 1 | 9 | 2 | 2 | 0 | 0 | 0 | 0 | 1 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 15 |
| NEM | 14 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 14 |
| NWM | 7 | 0 | 6 | 0 | 0 | 0 | 0 | 8 | 14 | 0 | 0 | 0 | 0 | 2 | 0 | 0 | 0 | 0 | 37 |
| SEM | 14 | 20 | 9 | 5 | 0 | 0 | 0 | 0 | 0 | 3 | 0 | 1 | 0 | 0 | 0 | 0 | 0 | 0 | 52 |
| SWM | 31 | 1 | 26 | 7 | 5 | 7 | 0 | 2 | 0 | 0 | 1 | 0 | 1 | 0 | 1 | 0 | 1 | 1 | 84 |
| SW | 0 | 0 | 16 | 3 | 0 | 0 | 4 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 1 | 0 | 0 | 24 |
| SE | 0 | 0 | 14 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 1 | 0 | 0 | 15 |
| Sum | 66 | 21 | 72 | 24 | 7 | 9 | 4 | 10 | 14 | 3 | 2 | 1 | 1 | 2 | 1 | 2 | 1 | 1 | 241 |
#1389. between θ츻 ([2013-02-14-1]) ڤӺε#1393. between Ū۷֤˭٤([2013-02-18-1]) ˰³Ƥꡥbetween Ūʰ۷֤ʬۤHelsinki Corpus ǤäĴƤߤĴη̡ѥ between η֤Ȥ 97 types, 793 tokens ǧ줿ʲϤ97ΰ۷֤֡Ǥ롥
be-twen, be-twene, be-twix, be-twyen, be-twyn, be-twyx, be-twyxe, betuen, betuene, betuh, betuih, betuixt, betun, betux, betuyx, betwe, between, betweene, betwen, betwenan, betwene, betweoh, betweohn, betweon, betweonan, betweonen, betweonon, betweonum, betweox, betweoxan, betwex, betwi, betwih, betwihn, betwinan, betwinum, betwioh, betwion, betwix, betwixe, betwixt, betwixte, betwixts, betwne, betwoex, betwonen, betwuh, betwux, betwuxn, betwyh, betwyn, betwynan, betwyne, betwyx, betwyxe, betwyxen, betwyxte, bi-tuine, bi-twen, bi-twene, bi-twenen, bi-tweohnen, bi-tweone, bi-tweonen, bi-twexst, bi-twext, bi-twihan, bi-twixst, bituen, bituene, bituhe, bituhen, bituhhe, bituhhen, bituien, bituih, bituin, bituix, bitunon, bitweies, bitwen, bitwene, bitwenen, bitwenenn, bitweon, bitweone, bitweonen, bitweonon, bitweonum, bitwex, bitwexe, bitwien, bitwih, bitwix, bitwixe, bitwixen, bitwyxe
793η֤δǤޤȤƽפΤưפǤϤʤϸʳˤ첻ΰ㤤̵뤹뤳Ȥˤ2ʹߤλҲʤȡ⤷и첻ˤμȤ߹碌ˤäƽפ㤨С"nm", "nn", "x", "xt" Ȥפϡ줾 betweonum, betweonan, betwyx, betwixt ʤɤη֤ɽ롥ʲɽϡHelsinki Corpus ˤʬȤεʤä O1 ʸűѸ1ˤλ10ˤ̻ŪѲΤǤ롥
| nm | nn | n | ne | x | xe | xn | xt | xte | xst | xts | h | hn | hnn | he | s | e | i | Sum | |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| O2 | 14 | 1 | 0 | 0 | 13 | 0 | 1 | 0 | 0 | 0 | 0 | 31 | 4 | 0 | 0 | 0 | 0 | 0 | 64 |
| O3 | 5 | 22 | 16 | 0 | 56 | 0 | 0 | 0 | 0 | 0 | 0 | 48 | 0 | 0 | 0 | 0 | 0 | 0 | 147 |
| O4 | 1 | 15 | 3 | 0 | 22 | 0 | 0 | 0 | 0 | 0 | 0 | 3 | 0 | 0 | 0 | 0 | 0 | 0 | 44 |
| M1 | 0 | 28 | 4 | 8 | 13 | 0 | 1 | 0 | 0 | 0 | 0 | 0 | 4 | 1 | 9 | 0 | 0 | 0 | 68 |
| M2 | 0 | 1 | 5 | 32 | 1 | 1 | 1 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 1 | 0 | 0 | 42 |
| M3 | 0 | 0 | 4 | 31 | 24 | 18 | 0 | 1 | 0 | 4 | 0 | 0 | 0 | 0 | 0 | 0 | 1 | 0 | 83 |
| M4 | 0 | 0 | 4 | 11 | 25 | 6 | 2 | 1 | 6 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 1 | 56 |
| E1 | 0 | 0 | 12 | 66 | 2 | 0 | 0 | 25 | 3 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 108 |
| E2 | 0 | 0 | 23 | 44 | 0 | 0 | 0 | 31 | 6 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 104 |
| E3 | 0 | 0 | 54 | 8 | 0 | 0 | 0 | 14 | 0 | 0 | 1 | 0 | 0 | 0 | 0 | 0 | 0 | 0 | 77 |
| Sum | 20 | 67 | 125 | 200 | 156 | 25 | 5 | 72 | 15 | 4 | 1 | 82 | 8 | 1 | 9 | 1 | 1 | 1 | 793 |
ε#1355. 20ꥹѸǽ̾ñפä ([2013-01-11-1]) Ǽ夲Bauer ν̾οΰפ˴ؤĴˤĤơҲ³롥The Times μΥѥˤ̻ŪĴ̤ơBauer ϷȴƺѤν̾Ǥ government 20δ֤ˡΰפ˴ؤƶ̣ʬۤȤȯ (Bauer 64--65)
20ˤϡgovernment ʣפ¿ΤΡ褫ŦƤȤꡤȤ館˱ñʣΤѰۤƤȤˤʤȡñʣפΰ㤤ؼоݤΰ㤤б褦ˤʤäʣȤѤȤˤϱѹܤؤñȤѤȤˤ¾ܤؤȤƤȤΤʸˡȤ⡤̣ʻؼоݡˤؤȰܹԤΤ褦ʲˡBauer (64) "Concord with government by meaning from The Times corpus, 1930--65" ΥǡƷǤ褦
| Year | British government | Non-British government | ||
|---|---|---|---|---|
| Singular | Plural | Singular | Plural | |
| 1930 | 3 | 15 | 12 | 3 |
| 1935 | 2 | 13 | 1 | 12 |
| 1940 | 2 | 14 | 4 | 2 |
| 1945 | 2 | 7 | 2 | 2 |
| 1950 | 1 | 26 | 26 | 0 |
| 1955 | 2 | 2 | 8 | 0 |
| 1960 | 0 | 23 | 8 | 0 |
| 1965 | 1 | 13 | 4 | 1 |
| Total | 13 | 113 | 65 | 18 |
ưοΰפˤĤƤϡ#930. a large number of people οΰס ([2011-11-13-1]) #1144. Ѹˤפ ([2012-06-14-1]) #1334. Ѹˤ̾ưοס ([2012-12-21-1]) εǰäƤ̤ˡѸˤ government team ʤɤν̾οΰפϡꥫѸǤϤäѤñǰפ뤬ꥹѸǤϤȤ館˱ñǤʣǤפȤ롥ΰ̲ϳͭΰפ˴ؤѰۤꥹѸˤĤƤߤȡ20̤ñפηޤäƤƤΤǤϤʤȤŦ롥Bauer (61--66) The Times μоݤȤѥҲ𤷤褦
Bauer ϡ1900--1985ǯ The Times corpus μ⤫ʤ륳ѥоݤˡ̾줬ñǰפΨBauer ϡĴ The Times μȤ˷ĥäʸΤˤĴǤꡤ줬ɬ⥤ꥹѸΤɽƤȤϤʤǤäạ̇̄դͿƤ롥ʲϡBauer (63) ΥդܸƤǿͤɤ߽ФƺΤǤ롥

ľȤƤʤ餻Сǯ0.3178%γǤȤʤäƤ롥ͤꤷʤȤ䥳ѥФʤɤͳˤꤳη̤ɤޤǿǤΤȤʤ뤬Bauer Ϻ٤ͿƤ餺ȽǤǤʤΤǤ롥
Bauer ϤˡѥǺǤ٤ι⤤̾ government üʿ뤳Ȥܤθ̾ˤĤơñפγƷΥդƱΤǡBauer (66) Υդ˴ŤƲΥդƺʤ餹ǯ0.1877%γǤǤ롥

η̤¿Dzη̤ˤȤɤޤ餶ʤ褦˻פ뤬ʤȤ⤵ĴʤƤ椯ΥˤϤʤ
ʤBauer 1930ǯˤη֤äȤ¤ˡꥫѸ줬ƶͿȹͤ뤳ȤϤǤʤȤƤ[2011-08-26-1]ε#851. ꥹѸФ륢ꥫѸαƶ2狼פȡˡ
government οΰפ˴ؤ뤪⤷ˤĤƤϡεǾҲ𤹤롥
Bauer, Laurie. Watching English Change: An Introduction to the Study of Linguistic Change in Standard Englishes in the Twentieth Century. Harlow: Longman, 1994.
ѸβäǤϡղõ䤬褯Ȥ롥ŪˤɤΤ餤褯ȤΤ⤽Ū˵ʸϤɤΤ餤٤ΤΤʤǡղõϤɤ줯餤γΤΤ褦ʵ顤ޤ٤ Biber et al. LGSWE Ǥ롥
ǽˤĤƤϡp. 211 ˲ͿƤ롥οˤƤĴCONV(ERSATION) Ǥ401ĵ䤬ޤޤƤȤåѥǤϡž̾塤䤬ȿǤƤǽ⤯ºݤˤϿͰʾ٤ǵʸƤϤǤ롥ƥȥפǤС礭 FICT(ION) ³NEWS ACAD(EMIC) Ǥϵʸ٤ϸ¤ʤ㤤
ˡƥ֥ѥˤơʸΤˤղõϤɤΤ餤p. 212 ˷ǺܤƤ̤ʲΤ褦ˤޤȤĤ100%ȤʤɽǤ롥
| (* = 5%; ~ = less than 2.5%) | CONV | FICT | NEWS | ACAD | |
|---|---|---|---|---|---|
| independent clause | wh-question | **** | ******* | ********* | ********** |
| yes/no-question | ***** | ***** | ******* | ******* | |
| alternative question | ~ | ~ | ~ | ~ | |
| declarative question | ** | * | ~ | ~ | |
| fragments | wh-question | * | ** | ** | * |
| other | *** | *** | * | * | |
| tag | positive | * | ~ | ~ | ~ |
| negative | **** | * | ~ | ~ | |
Tag questions (i.e., regular questioning expressions tagged onto a sentence) exist in both American and British English, with British speakers perhaps using them more than Americans: "That's not very nice, is it?" Peremptory and aggressive tags tend to be used more in British English than in American English: "Well, I don't know, do I?" (192)
ǰʤ顤Biber et al. Ǥղõ٤αƺΤ뤳ȤϤǤʤӡƤΥѥĴ٤ɬפ
Schmitt, Norbert, and Richard Marsden. Why Is English Like That? Ann Arbor, Mich.: U of Michigan P, 2006.
Biber, Douglas, Stig Johansson, Geoffrey Leech, Susan Conrad, and Edward Finegan. Longman Grammar of Spoken and Written English. Harlow: Pearson Education, 1999.
Cheshire (115) ɤǤơѸ˴ؤ뵭ҤȤơäˤ¿ȤȤڤľŪˤϳΤˤΤ褦˻פ뤬ҴŪդϤΤȡLGSWE äƤߤȡϢ뵭Ҥ pp. 159--60 ˸Ĥä
ˤ͡ʼब뤬4ĤλѰΤ줾ˤĤơѥѤ "Distribution of not/n't v. other negative forms" Ĵ̤Ƥ100դɽǼ

| not/n't | other negative forms | |
|---|---|---|
| CONV | 19500 | 2500 |
| FICT | 9500 | 4000 |
| NEWS | 4500 | 2000 |
| ACAD | 3500 | 1500 |
Helsinki Corpus (The Diachronic Part of the Helsinki Corpus of English Texts) 1991ǯ˸ư衤Ѹ˥ѥθĤȤƽѤƤHC ϸߤǤƤ餺ܥ֥Ǥ#381. oft often ʬ̻ۤŪѲ ([2010-05-13-1]) Ϥᡤhc γƵǸڤƤ
HC ܳŪ˻ȤʤˤϡΥޥ˥奢ɤɬפ롥Ȥ櫓̥֥ѥθϲƤɬפ뤷COCOA Format ˤ뻲ȥפCOCOA Format ϡHC ΥƥˤΥƥȤ˴ؤξͿ뤿ηǤ롥ƥƥȤˤĤơǯ塤Ԥ̡ʸʸʤɤξηˤͿƤ롥ѼԤϡξѤ뤳ȤˤꡤξƥȤӽФȤǤȤ櫓
HC COCOA Ѥιʤߤؤˤ뤿ˡޤɽˤޤȤǡ١ (SQLite)
A = "author"
B = "name of text file"
C = "part of corpus"
D = "dialect"
E = "participant relationship"
F = "foreign original"
G = "relationship to foreign original"
H = "social rank of author"
I = "setting"
J = "interaction"
K = "contemporaneity"
M = "date of manuscript"
N = "name of text"
O = "date of original"
P = "page"
Q = "text identifier"
R = "record"
S = "sample"
T = "text type"
U = "audience description"
V = "verse" or "prose"
W = "relationship to spoken language"
X = "sex of author"
Y = "age of author"
Z = "prototypical text category"
ŵŪʸȤƵƤ
# ɽΤƸ[ | ѥڡ ]
select * from hccocoa
# ʬ̤Υƥȿ
select C, count(*) from hccocoa group by C
# ƥȥ̤Υƥȿ
select T, count(*) from hccocoa group by T
# ME ˻ʬƤƥȤγƼ
select B, C, D, V from hccocoa where C like 'M%' order by C
ε#1321. BNC Frequency Extractor ([2012-12-08-1]) ˰³ANC (American National Corpus) ˴ŤɽANC Second Release Frequency Data Υڡ˸ƤΤǡ"ANC Frequency Extractor"
# եƥȤǡƺȤ "diarrhoea" vs. "diarrhea" ֻ٤ǧ
select * from written where word like "diarrh%"
# եƥȤǡƺȤ "judgement" vs. "judgment" ֻ٤ǧʤ¾[2009-12-27-1]ε#244. ֻαƺΥꥹȡפֻǤ椯Ȥ⤷
select * from written where word like "judg%ment%"
# -ly ǽʤõflat adverb ⤷ʤõ
select * from anc where lemma not like "%ly" and pos like "RB%"
# -s ǽõadverbial genitive ̾Ĥ⤷ʤõ
select * from anc where pos like "RB%" and word like "%s"
# ñ̾ʣ̾ token Ӥ written subcorpus spoken subcorpus ǡ[2011-06-07-1]ε#771. ̾ñʣ١פȡ
select pos, sum(freq) from written where pos in ("NN", "NNS") group by pos
select pos, sum(freq) from spoken where pos in ("NN", "NNS") group by pos
select pos, sum(freq) from anc where pos in ("NN", "NNS") group by pos
ANC ͭȴ褵줿 OANC (Open American National Corpus) ̵ANC ڤ OANC ˤĤƤϡ#708. Frequency Sorter CGI ([2011-04-05-1]) #509. Dracula ˸ whilst (2) ([2010-09-18-1]) ȡ
"BNC Frequency Extractor" "ANC Frequency Extractor" Ȥ߹碌ƻȤСäαƺˤĤ٤δñĴǤ롥
Adam Kilgarriff Ƥ BNC database and word frequency lists 顤Ф첽Ƥʤɽ (unlemmatised lists) ɤǤ褦˥ǡ١館
# եƥȤǡƺȤ "diarrhoea" vs. "diarrhea" ֻ٤ǧ
select * from written where word like "diarrh%"
# s ǻϤޤʬι⤤
select * from variances where word like "s%" order by variance desc limit 100
# 첻Ѱۤʣñ١cf. #708. Frequency Sorter CGI([2011-04-05-1]) Ǥ lemma ä
select * from bnc where word in ("foot", "goose", "louse", "man", "mouse", "tooth", "woman") and pos = "nn1" order by freq desc
# 첻Ѱۤʣ
select * from bnc where word in ("feet", "geese", "lice", "men", "mice", "teeth", "women") and pos = "nn2"
# POSǤޤȤ٤ι⤤ˡä 'demog'
select pos, sum(freq) from demog group by pos order by sum(freq) desc
# Ǥ¿Ȥ̾
select * from variances where pos like "n%" order by variance desc limit 100
# Ǥ¿Ȥƻ
select * from variances where pos like "aj%" order by variance desc limit 100
ʤФ첽Ƥɽ (lemmatised list) ˤĤƤϡ٤ˤ800ʾ帽롤6318̤ޤǤθФΤߤ˸ꤵƤꡤθġϡ#708. Frequency Sorter CGI ([2011-04-05-1]) ȤƼƤ롥Ϣơ#956. COCA N-Gram Search ([2011-12-09-1]) ⻲ȡ
ѸˤϡǾ most mest Ȥ첻ȼäƸ뤳ȤʤʤѸʹߡԤѤƤäξεʬϤɤˤΤ
most Proto-Germanic *maistaz ̤뤳ȤǤޥǤ Du. meest, G meist, ON mestr, Goth. maists ʤɤʸڤ롥§˽СűѸ māst ȤʤϤǤꡤºݤˤη֤ Northumbrian dzǧΤΡǤϳǧʤǤϡ첻ȼ West-Saxon mǣst Kentish mēst Ѥ줿OED ˤС첻ϡlǣst "least" ȤȤ롥첻ηȤ mest(e) Ȥ֤ѸؤѾ졤Ǥ15ޤǻȤ줿
˵ķ֤ϡѸǤϸ첻ηȯãȤ most(e) Ȥ֤¿Ѥ줿Ǥ̲ȤλΰŪʿ˲äӵ mo, more 첻ȤäΤǤϤʤ롥
ŪˡѸʹߤˤϥޥĸ줫ε§Ūȯã most ɸŪȤʤäƤ椭űѸ줫ѸˤѤ줿 mest ɸफϼƤäְΡפ̣Ѹ formest (cf. ӵ former) 15 foremost ȤƺʬϤ줿طʤˤϡҤ most ˤ mest ִͿƤ뤫⤷ʤäȤ⡤űѸꡤǾ -est Τ -ost Ȥ褯Ʊ줿ΤǤꡤǾ˴ؤˤơξ첻θؤϾˤȤʤΤ⤷ʤ
ʤPPCME2 Ǥäȸ첻 (ex. most) 첻 (ex. mest) ʬۤĴ٤ƤߤȡԤ354㡤Ԥ168ҥåȤHelsinki Corpus ǤñĴѸǤ⸽ɸѸƱͤ most ήäȤϴְ㤤ʤ褦
[2010-12-25-1]ε#607. Google Books Ngram ViewerפǾҲ𤷤 Google Υѥġˡ쥿դ줿եǤ Google Books Ngram Viewer θѤʤɸĤθϤǤ褦ˤʤäξҲˡϡSyntactic Annotations for the Google Books Ngram Corpus ǻȤǤ롥
ߡGoogle Books Ngram Corpus English, Spanish, French, German, Russian, Italian, Chinese, Hebrew 8ΥѥޤबѸ쥳ѥ˴ؤ¤ꡤ4,541,627ʬ468,491,999,492 tokens ʤĶƥȡǡ١ȤʤäƤ롥ǡåȤǽ
줿쥿ϡŪˤСʻ (POS) Ƚط (head-modifier) Ǥ롥ɸդ׳Ū˼ưǹԤʤƤ롥ʻϰʲ12ब̤롥
NOUN (nouns), VERB (verbs), ADJ (adjectives), ADV (adverbs), PRON (pronouns), DET (determiners and articles), ADP (prepositions and postpositions), NUM (numerals), CONJ (conjunctions), PRT (particles), '.' (punctuation marks), X (a catch-all for other categories such as abbreviations or foreign words)
ϼȤƤϡ㤨 "burnt" Τ褦˸뤳ȤǤ뤷"burnt_VERB" Τ褦ʻꤷ뤳ȤǤ롥 3-grams ϢǤ "_ADJ_" Τ褦ʰѤǤ롥ʾΥѥ碌ơ"the _ADJ_ girl_NOUN" ʤɤǽطλǤϡ"hair=>black", "read=>book" ʤɤϤǤ䤽¾ΥΥȤʤǤϤȤǽȤʤäƤ롥
̾ưˡͭƤˤĤơʻ̤Ѳߤͤ褦travel ̾ǤưǤ⤢뤬Ѹ쥳ѥΤоݤȤˤС20ä̾ˡưˡɤȴȤ狼롥оݥѥꥫѸꥹѸڤؤӤȡԤ̾줬ư٤ξɤȴΤ1960ǯȤä٤
ۤˡhave a look ڤ take a look ȤɽγĴ٤褦Ȥˡ괧θ˷ƻʤɤǽθ"have>=look, take>=look" ʤɤȸƤߤꥫѸǤ take Ѥɽ1970ǯɤȴƤ뤬ꥹѸǤ20˽˳礳Ƥ뤬ޤ have ѤɽɤĤƤʤ
[2010-03-04-1]ε#311. girl Ȥ褯 collocate ƻϲפǡȸζ (collocation) ¬ˡ (association measure) ˤϤĤμब뤳ȤѥؤǤϡLog-Likelihood Test ȤˤˡŪ褯ȤƤ뤬줾ηˡˤħΤǡʤ٤ʣˡΤ褤[2010-03-04-1]ƤȽʣʬ⤢뤬BNCweb ǼƤ7ηˡγơˤĤ Hoffmann et al. (149--58) Ȥʤ顤ħѤΥҥȤ
Ƽηˡϡ(a) (frequency of co-occurrence)(b) ͭ (significance of co-occurrence)(c) եȡ (effect-size) 1ġ뤤ʣȤ߹碌˴ŤƤ롥(b) ϡŪͭդǤȤγο٤ɽ魯ɸǤꡤζɽ魯ΤǤϤʤȤդɬפ롥(c) ϡѻ٤ȴ٤ȤδܤȤɸǤ롥
(1) Rank by frequency
ѻ붦٤ΤΤѤ롤ǤñľŪʼ١¾ηˡΤ褦ʣϤۤɤƤ餺ɸȤƤϺǤƤǽɵʤɤ̤뤳Ȥ¿̾ζʬϤˤѤʤ
(2) Log-likelihood
ͭѤ롥BNCweb ΥǥեȤηˡǡѥǹѤƤ롥ǽɵʤɤζˤƹ٤θȤζ䡤դ˶ˤ٤θ1, 2ʤɡˤȤζϤ롥٤ι⤤Ȥ߹碌˹ͿȤħꡤˤդפ롥
(3) Mutual information (MI)
եȡѤ롥ˤ褯ѤƤˡѤäƤ¿դפ롥ǽɵʤɤȤΤդ줿ŪӽƤϤ褤ȿ̡٤ζɽؤФ꤬㤷Фαƶ뤿ˡBNCweb Ǥ "Freq(node, collocate) at least" 10ʾꤹ뤳Ȥ侩롥ˤꡤ"conspicuous and intuitively appealing collocations involving words of intermediate frequency" (Hoffmann et al. 154) ⤭ĦȤʤ롥
(4) T-score
٤ȶͭθˡ٤1ʲ٤εʶɽˤĤƤ Rank by frequency Ȼ褦ʿ٤ι⤤ɽˤĤƤ϶ͭȿǤ롥ޤѻ٤٤ɬ⤯ʤ롥Log-likelihood ̤Ȥʤ뤳Ȥ¿٤ؤΥХϰضʤ롥ΡɤΤΤ1000礭ˡ̤ȯ뤳Ȥ롥
(5) Z-score
ͭȥեȡθˡ٤ζɽˤϥեȡŻ뤹뤬٤ζɽˤϤޤǥեȡ˴꤫ʤLog-likelihood MI ξħ褦ʡХμ줿ɸǤ롥MI Ʊͤˡ٤ζɽؤΥХߤΤǡ"Freq(node, collocate) at least" 5٤ꤹΤ褤Ȥ롥
(6) MI3
٤ȥեȡθˡMI ΤɽؤнŤ٤Ƥ롥ٶɽˤϥեȡٶɽˤ϶٤Ū褯ȿǤ롥POS ˤȤȤѤȸŪʣ줫ʤѸʤɤμФ˰Ϥȯ롥ΤȤƤϹٶɽؤΥХŪʶʬϤˤϸʤ
(7) Dice coefficient
MI3 Ʊͤˡ٤ȥեȡθˡMI3Ȱۤʤꡤٶɽˤ϶٤ٶɽˤϥեȡ褯ȿǤ졤ξԤڤؤޤʤΤħŪǤ롥ڤؤϡΡɤΤΤ٤ɽ٤10ܤۤɤǵȤ롥иŪˡZ-score Ȼ褦ʷ̤뤬Z-score ۤ٤˴ŤХʤ
ʾΤ褦¿ढäܰܤꤹ뤬Hoffmann et al. θˤСñηˡȤƤ Log-likelihood MI ǡηˡȤƤ Z-score Dice ȤΤȤǤ롥
͡ʷˡˤĤƤϡAssociation measures ȡ
Hoffmann, Sebastian, Stefan Evert, Nicholas Smith, David Lee, and Ylva Berglund Prytz. Corpus Linguistics with BNCweb : A Practical Guide. Frankfurt am Main: Peter Lang, 2008.
Powered by WinChalow1.0rc4 based on chalow