hellogѸ˥֥     ChangeLog ǿ     ƥǿ     ڡ 1 2 3 4 5 6 7 8 9 10 11 12 ڡ / page 10 (12)

corpus - hellogѸ˥֥

ǽ: 2026-07-15 01:27

2011-04-28 Thu

#731. -dom ̻̾Ūʬ [suffix][oed][corpus][productivity]

[2009-05-18-1]ε-dom ̾פǤϸѸǻȤ -dom 򤤤Ĥ󤲤̻Ūʴ餳įƤߤBauer (220) ˤȡ-dom ϰ٤λȤߤʤۤɤ˿षƤѸǤϰᤷƤƤȤ

-dom    This suffix forms abstract, uncountable nouns from concrete, countable ones. For a long time it was thought that the suffix was moribund or totally non-productive, but Wentworth (1941) showed that it had never completely died out, and it is still productive in contemporary English, though not very much so. Recent examples include Dollardom, fagdom, gangsterdom, girldom (all OEDS). (220)


-dom ϸ§Ȥ̾δΤղä̾뤬freedom Τ褦˷ƻδΤղä⤢롥
OED ̻ŪʬۤĴ٤Ƥߤ[2011-01-05-1]ǾҲ𤷤OED θ̤äȤʬह CGIפѤȤ -dom 夲ʲΤ褦˻вSodom ʤɤλ¿ϺäƤꡤäȸƵդΤϺ绨Ĥʿ夲Ȥ򤵤줿ͥǡϤΥڡHTML򻲾ȡ

Diachronic Distribution of -dom Words by OED

Ѹ줫ѸˤƤΤ露Ȥ衤19ȯϰŪǤ롥20βФϡ¤ȿǤƤΤ뤤 OED θüλˤΤˤƤ19ʹߤο -dom äϤ٤Ƥٸǡnonce-word ¿Frequency Sorter ˤȡANC (American National Corpus) 10ʾѤƤΤϡfandom, boredom, stardom, fiefdom
ܼ (productivity) Ū˷׻Τ񤷤Ȥ (Baayen and Lieber) -dom 19ȯ2021ˤɤ³ƤΤľŪªܼȤϵҴŪˤɤΤ褦˵ҤΤ˥ѥɤΤ褦˳ѤǤΤ-dom ܤǤ⡤꤬͡夬äƤ롥

Bauer, Laurie. English Word-Formation. Cambridge: CUP, 1983.
Baayen, Harald and Rochelle Lieber. "Productivity and English Derivation: A Corpus-Based Study." Linguistics 29 (1991): 801--43.

[ | ѥڡ ]

2011-04-08 Fri

#711. Log-Likelihood Tester CGI, Ver. 2 [corpus][bnc][statistics][web_service][cgi][lltest]

ʲˡѤ Log-Likelihood Tester, Ver. 2 ʸ褦ˡϥǡΥեޥåȤ䡤⡼ɤŬڤ򤵤ƤʤˤϥСǥ顼ǽΤա

each-line mode lump mode


[2011-03-25-1]εǡѥǤ褯Ѥпٸ ( Log-Likelihood Test ) η׻ Log-Likelihood Tester, Ver. 1 Ver. 1 ϡѥ̣ʤ2ĤΥѥǤΥɡʷˤνи٤١ѥ֤κͭդǤ뤫ɤꤹΤä
Log-Likelihood Test ϾҤŪѤ뤳Ȥ¿ȻפVer. 1 ǤϤƵǽòΤŪʣԡʣʬɽͿǡбпٸԤʤ⤢롥㤨Сε[2011-04-07-1]ǡѸˤ though although νиˤĤ BNC ˴ŤĴҲ𤷤Text Domain ȤΨϡξδ֤ŪˤɤٰפƤ롤뤤ϰפƤʤȤߤʤȤǤΤΥդ顤although ϳؽѻʸ¿though Ϻʸ¿ȤľŪʡְפŪˤϤɤΤ褦ɽΤ
Τ褦ʾˤϡΤ褦ɽͤ100νи٤ɸಽѤߡˤ򥳥ԡϥܥåŽդ롥"lump mode" ˥åؤ"Go!" 롥ʥǥեȤ "each-line mode" ǡ Ver. 1 ƱΥ⡼ɡ

    thoughalthough
Natural and pure sciences56.380.13
Applied science37.3668.31
World affairs45.8168.2
Social science48.9863.38
Commerce and finance46.1857.21
Arts74.0752.93
Leisure45.8549.46
Belief and thought70.7846.75
Imaginative prose80.226.37


̤ϡ1ԤɽȤƽϤ롥though although ɽ魯2οͤ¤ӤŪˤɤΤ餤Ƥ뤫׻Ƥ롥ȤƤϡξ Text Domain Ȥ٤¤Ӥκ p < 0.0001 Ȥ˹⤤٥ͭդǤꡤξνи Text Domain ˤäƤۤܳμ¤˰ۤʤȤ롥
ϥܥåǡν񼰤ϡֶڤʬɽɽƬɽ¦ϤάġץΤ褦ɽƬɽ¦ξޤˤϡΥ϶ˤƤɬפꡥ
"each-line mode" εǽ Ver. 1 ȸߴʤΤǡϷ⤽򻲾ȡ Ver. 2 "each-line mode" ǤϡϷ̤򥷥ץˤƤʵդˡܤ׻ͤˤ Ver. 1 Τۤͭѡˡ
Log-Likelihood Test γפˤĤƤϡ[2011-03-24-1]ε򻲾ȡ

Referrer (Inside): [2012-10-26-1]

[ | ѥڡ ]

2011-04-07 Thu

#710. though although θˡκ (2) [bnc][corpus][lltest][conjunction][statistics]

ε[2011-04-06-1]ǡthough although θˡκ˿줿Ʊǡ
4000Ķʤ The Longman Spoken and Written English Corpus (the LSWE Corpus) ȤѸʸˡBiber et al. (845--46) ǤϼΤ褦ˤ롥

Both of these subordinators [though and although] occur in all four registers [conversation, fiction, news, and academic prose], although the registers show different preferences of use. Conversation and fiction show a slightly greater use of though (concessive clauses are, however, uncommon in conversation generally). News shows no particular preference. In academic prose, although is about three times as frequent as though. Although seems to have a slightly more formal tone to it, fitting the style of academic prose . . . . The greater use of although by writers of academic prose may also result from an attempt to distinguish this subordinator from the common use of though as a linking adverbial in conversation . . . .


ޤƱ p. 842 ɽϡŪ though fiction ¿although academic prose ¿Ȥǧ롥ˤ뺹ƤȤη̤
Τ褦Ըơ BNC ( The British National Corpus ) ˤꤳΤƤߤ롥BNCweb ǡ{although/CONJ}, {though/CONJ} 򤽤줾측Written/Spoken, Text Domain, Sex of Author/Speaker, Perceived Level of Difficulty ʤ͡ʥѥ᡼ǽиʬۤʬϤΩä̤ʲ˼ʿͥǡϤΥڡHTML򻲾ȡˡ
ޤWritten/Spoken κˤĤƤϡͽۤȤꡤξȤ Written ؤФ꤬㤷ʺ۷ though 0.66344 although 0.49770 ǡ餫˽񤭸դФˡLog-Likelihood Test Ǥϡp < 0.0001 Υ٥ǽ񤭸դäդͭպΤ˼줿
񤭼ꡤäˤ뺹ⶽ̣񤭸դäդξǡalthough ͭպäλѤФäƤ롥though ˤĤƤϡ although ۤɸǤϤʤʤ񤭸դǤ p < 0.05 ͭպˡ
ˡText Domain ̤٤ߤ롥9 Text Domain ̤ ( Natural and pure sciences, Applied science, World affairs, Social science, Commerce and finance, Arts, Leisure, Belief and thought, Imaginative prose ) 100νиɸಽͤǡξ Text Domain ٤򥰥ղΤʲοޤ



Text Domain ˤäξνи٤оŪʷ뤳Ȥ狼롥Ū sciences ( = academic prose ) although ΩImag(inative) Prose ( = fiction ) though ¿Log-Likelihood Test ǤϡText Domain ˤиκ p < 0.0001 ͭդǤ롥
ľŪˤԸη̤ͽۤȤǤϤ뤬although ν񤭼ˤؽѻʸǸѤȤ޼줿

Biber, Douglas, Stig Johansson, Geoffrey Leech, Susan Conrad, and Edward Finegan. Longman Grammar of Spoken and Written English. Harlow: Pearson Education, 1999.

Referrer (Inside): [2011-04-10-1] [2011-04-08-1]

[ | ѥڡ ]

2011-04-05 Tue

#708. Frequency Sorter CGI [corpus][bnc][statistics][web_service][cgi][lexicology][plural]

餫δǽ᤿ñΥꥹȤ򡤰Ū٤ν¤ؤȤ롥㤨С[2011-03-22-1]褦ˡ٤Ե§ʿ񤤤ȤδطĴ٤ȤˡܤʷˤΰŪ٤Τɬפ롥Ūˤϡ[2010-03-01-1]ǾҲ𤷤褦絬Ϥѥѥ˴ŤɽͭѤǤ롥BNC lemma-pos list (122KB) ANC word-tagset list (7.2MB) ʤɤθĤҤȤĸٿٽ̤Ĵ٤Ƥ椱Ф褤¿ˤݤǡ嵭2Ĥɽ顤Ϥʷˤ٤Ƚ̤Ф CGI
ԤǤ⥹ڡǤ⥫ޤǤ褤Τڤ줿ñꥹȤʲΥܥåϤ"Frequency Sort Go!" 򥯥å롥Ϸ̤ٽ̤ι⤤˥Ȥˤϡ"sort by rank?" 򥪥ˤʥǥեȤǥ󡥥դˤȡϽ˽Ϥˡ㤨СɸѸ˻Ĥ i-mutation 򼨤ʣϰʲ7ΤߤǤʣ졤ʣ[2011-04-01-1]ˤ sister(e)n Ͻˡ򥳥ԡƥܥåϤ롥

foot, goose, louse, man, mouse, tooth, woman


     sort by rank?


ޤBNC lemma-pos list ˤϤɽ1 BNC Τ顤٤ˤ800ʾ帽롤6318̤ޤǤθФ ( lemma ) ϿƤ롥äơ٤β goose, louse ˤĤƤ϶ȤʤäƤ롥٤Ե§شطͤݤ˻ͤˤʤ
ˡANC word-tagset list ˤϤ³ɽ BNC ΤΤ⵬Ϥ礭Ĥ٤22,164,985ͭ ANC (American National Corpus)Penn Treebank Tagset ˤäƥ饹Ϳ줿ñ̤Ǹ󤵤줿ꥹȤǤ롥åȤ٤ΤɤߤˤưͿ˵륨顼⾯ʤ餺ޤޤƤ뤬BNC ΤΤ٤θʷˤϿƤΤǡgoose louse پ⸽롥ɽǤ WORD FORM Ȥ٤ǧǤ뤿ᡤľ geese lice ٤Τ롥
Frequency Sorter ӤȤꤷƤΤϡ嵭Ե§ʣ򼨤췲ʤɤ٤Ƚ̤ΰĴä¾ˤӤϤ뤫⤷ʤʲˡפĤ⡥

1ñ줫ȤΤǡlike Τ褦¿ʻϤơʻʤ뤤ϥͿ줿饹ˤȤ٤Ф롥
ҥåȿǧˤϡѥΩ夲ɬפʤ
ʸץ쥼ǡŪǽ᤿ɴñꥹȤ椫ŵŪ㡤ʬ䤹10ĤۤɼȤʤɡ٤ι⤤10Ĥ٤Ф褤㤨С[2011-03-29-1]󤷤 sur- ƬˤñꥹȤΤ㼨˺Ǥդ路10Ĥ֤ʤɤŪˡ٤˴Ť֤Τۤ䥢ե٥åȽڤʤȤ¿ʺ塤ܥ֥ɮ˳Ѥͽˡ
Ƥ줾ɽŪʥѥ˴ŤɽѤƤΤǡֻ֤ʤɤ٤αƺǧΤ˻Ȥ롥
ʼºݤˤ lemmatisation ɬפŬʱʸǤߤơ̯٤㤤줬ޤޤƤʤĴ٤롥٤ΥġʤΤǡ¾顦ؽŪˤȻȤ뤫⤷ʤ

[ | ѥڡ ]

2011-04-01 Fri

#704. brethren and sister(e)n [plural][analogy][ame][i-mutation][relationship_noun][corpus][coca][coha]

ε[2011-03-31-1]ǡűѸο²̾ζɽ򸫤brethren εˤĤƤڤȴϢƿ²̾줪դ ( analogy ) ⤦ĵ󤲤褦brethren Ȥ sister(e)n Ȥʣ롥MED εҤˤ褦ˡѸǤ -(e)n Ϥ̤Ǥꡤ-s ̲Τ brother ξƱʹߤǤ롥դϻΰʤΤǡܺ٤ʥǡäƤ롥ѸǤ⥤󥰥ɤǤ -s ͥǤϤλ sister ʣϸ§Ȥ -n 뤤첻θݤƤ뤳Ȥϴְ㤤ʤ ( Hotta, p. 256 )
ơsister(e)n ϸѸĤäƤ뤬brethren Ȱۤʤꡤ̾MˤϵܤƤʤBNC ( The British National Corpus ) ǤҥåȤʤäCOCA ( Corpus of Contemporary American English ), COHA ( Corpus of Historical American English ) ǤϤ줾4㡤1519ȾʹߤˤҥåȤäѤ饢ꥫѸʹ뤳Ȥʬ롥COCA 1ĵ󤲤롥Ƥ "CNN Crossfire" ǤֻϰѼԡˡ

Well, you know, I hate to correct you, but you made the same mistake many of your liberal brethren and sisteren, have said in analyzing this dissent by Judge Stevens.


COCA, COHA ξѥη19Τ16ޤǤ brethren and sister(e)n ȤƸ졤˥եѤ졤dear my ԤƤӤλȤ¿brethren Ʊͤ˽ŪȹŪʸ̮ǸƤ褦ꤵ줿ȤƤΤۤʸŪʸ̤⤢Τ⤷ʤϢơOED sister θ5ѤƤ"In the vocative, as a mode of address, chiefly in transferred senses. Also colloq. as a mode of address to an unrelated woman, esp. one whose name is not known."
äѤ饢ꥫѸѤ뤳ȤˤĤƤϡMencken (502) Ƥ롥

Sisteren or sistern, now confined to the Christians, white and black, of the Get-Right-with-God country, was common in Middle English and is just as respectable, etymologically speaking, as brethren.


sister(e)n Ȥʣ˴ؤŪϡḽ奢ꥫѸǤλѤѸη³ȤƤȤ館٤뤤ϥꥫѸDzƤ⤿餵줿ȤƤȤ館٤Ǥ롥OED ˤȡsister(e)n ϰŪʸϸȤƤ16ȾФѤ줿Ȥ롥Ѹ䥤ꥹѸޤ᤿Ĵʤʬʤ(1) brethren Ȥϻ鷺ꤽǤ뤳ȡ(2) brethren ȵӱƧΤǸƤӤʤɸä˹ޤ줽Ǥ뤳ȡ2饢ꥫѸǤκƷȹͤΤǤϤʤѸ츻Ū sister(e)n Ф줿餤顤ѸDzƺ줿ȤƤԻ׵ĤϤʤ
sister(e)n ̾μˤϺܤäƤʤ餤Υ쥢ʣbrethren, children, oxen (but see [2010-08-22-1]) Ʊ˻Ĥ뾯 -en ʣ֤Ƥ롥

Hotta, Ryuichi. The Development of the Nominal Plural Forms in Early Middle English. Hituzi Linguistics in English 10. Tokyo: Hituzi Syobo, 2009.
Mencken, H. L. The American Language. Abridged ed. New York: Knopf, 1963.

Referrer (Inside): [2011-04-05-1]

[ | ѥڡ ]

2011-03-25 Fri

#697. Log-Likelihood Tester CGI [corpus][bnc][statistics][web_service][cgi][lltest][sociolinguistics]

ε[2011-03-24-1] Log-Likelihood Test ˤ׻ˤ Rayson Log-likelihood calculator ѤФ褤ȽҤ٤ºݤθκݤ˺Ȥ⤦ưȻפäΤ CGI 򼫺Ƥߤ٤ϤȻפȤꤢ



ΥƥȥܥåϤ٤ǡϡֶڤɽη1ܡʾάġˤϥѥ̾2ܰʹߤϥɤȴѻٿʥҥåȿˡǽԤϳƥѥΥʸˡ"#" ǻϤޤԤϥȹԤȤ̵뤵롥1ܤΥϾάġ
ʲΥƥȤϥץ롥[2010-09-11-1]εǼ夲ƥӹѤƻӵȺǾޤ˥ȥå20٤BNCweb äե֥ѥüԤ̤ɽǤ롥ΤޤޥԡϥܥåŽդȡϷ̤ǧǤ롥

    BNC_Male_SpeakersBNC_Female_Speakers
new14991
good408310
free17375
fresh84118
delicious1234
full210107
sure532328
clean197223
wonderful270258
special17782
crisp1016
fine347215
big470415
great20396
real16380
easy326157
bright113110
extra347203
safe18292
rich12045
#--------
corpus_size49499383290569


˽֤ͭպä礭ΤϡбԤ֤ɤĤ֤줿 fresh, delicious, clean, wonderful, big ǡٿ˴ŤƷ׻줿 Diff_Co ( "Difference Coefficient" ֺ۷ ) ޥʥǤ뤳Ȥ顤ħŪʷƻȤȤˤʤ롥big ϰճʵ⤷̤Ǥ롥Фäͭպ򼨤ΤϲǼ easy rich Ǥ롥η̤Ϥɤ߹ळȤǤܺ٤Ĵ٤뤳ȤǤ롥ηƻȤϡüԤǤϤʤʹ̡ǯ𡤼Ҳ񳬵ʤɤ򼴤ĴƤ⤪⤷ȱѤǤ롥

Referrer (Inside): [2011-04-08-1]

[ | ѥڡ ]

2011-03-24 Thu

#696. Log-Likelihood Test [corpus][bnc][statistics][lltest]

[2010-03-04-1]εǿ줿ѥؤǤϳƼ׼ˡѤ롥ĤˡΤʤǤ⡤ɽΥѥ֤٤Ӥꡤcollocation ٹ礤¬Τ˹ѤƤΤ Log-Likelihood Test ( LL Test, G Test, G2 Test ʤɤȤ˸ƤФ븡Ǥ롥ѥθ줿ʤΤǥΰۤʤ륳ѥ֤ǤӤǽǤꡤƱŪǰˤ褯ѤƤ2踡 ( Chi-Squared Test ) ⤤ĤǤ줿ˡɾƤꡤǶΥѥǤϹѤƤ롥㤨С2踡ϴ٤5꾯ʤȤٸ򰷤Ȥѥ礭ΤȾΤӤȤ˿㤯ʤ뤬Log-Likelihood Test Ϥαƶˤ [ Rayson and Garside 2 ]
Log-Likelihood Test δŪʹͤϡѥȤˤɽδԤи١ʴ١ˤФͤȼºݤ˽и١ʴѻ١ˤκñʸȹͤۤɤ˶Ƥ뤫ɤȽꤹȤΤǤ롥ȤơΤ褦ʥǥBNC ( The British National Corpus ) äե֥ѥȽ񤭸ե֥ѥ̤ξ֥ѥ֤ f*ck Ȥ four-letter word ٤Ӥ롥BNCweb ꤳΥɤ򸡺ȡΤ褦ʷ̤줿

CategoryNo. of wordsNo. of hitsDispersion (over files)Frequency per million words
Spoken10,409,85857963/90855.62
Written87,903,571743172/3,1408.45
total98,313,4291,322235/4,04813.45


׽ۤɤޤǤʤDZ "Frequency per million words" 򸫤Сf*ck Ūäդ¿Ѥ뤳Ȥʬ뤬ϤŪ΢դ롥ޤ̵Ȥơäե֥ѥȽ񤭸ե֥ѥδ֤Ǥ f*ck ٺϸϰǤꡤθ˴ؤξԤ˰̣Τ뺹Ϥʤפꤹ롥Ωϡäե֥ѥȽ񤭸ե֥ѥδ֤Ǥ f*ck ٺϸϰǤʤθ˴ؤξԤκϰ̣פȤʤ롥̵⤬ٻ뤫ɤΤŪǤ롥

 Corpus 1Corpus 2Total
Frequency of wordaba+b
Frequency of other wordsc-ad-bc+d-a-b
Totalcdc+d


Log-Likelihood Test Ѥ Log-Likelihood ratio пפϡɽΤdzƥ֥ѥ ( c, d ) ȡƥ֥ѥǤ f*ck ٿ ( a, b ) ʬɽˤޤȤ᤿ǡ줾δ E1 E2 򲼤 (1) μǵᡤͤ (2) μƵ롥

(1) E1 = c*(a+b)/(c+d); E2 = d*(a+b)/(c+d)
(2) LL = 2*((a*log(a/E1))+(b*log(b/E2)))

f*ck οͤǷ׻ȡʲΤ褦ˤʤ롥

E1 = 10409858*(579+743)/(10409858+87903571) = 139.979170861796
E2 = 87903571*(579+743)/(10409858+87903571) = 1182.0208291382
LL = 2*((579*log(579/139.979170861796))+(743*log(743/1182.0208291382))) = 954.2115

Log-likelihood ratio Ȥ 954.2115 ȤͤФ롥ˤͤŬڤͭտ̾ 5%, 1%, 0.1%ˤб륫ͤӤ롥2 * 2 ʬɽФ׻Ǥϼͳ1ΥͤѤ뤳ȤˤʤäƤꡤͤͭտ 5%, 1%, 0.1% νˤ줾 3.84, 6.63, 10.83 Ǥ롥954.2115 Log-Likelihood ratio ͭտ 0.1% б 10.83 ⤺äȹ⤤Τǡ0.1% ͭտǵ̵ϴѤ롥СŪˤϵ̵⤬ǤΨ 0.1% ˤޤȹͤƤ褤ȤȤǤ롥Τ褦ˤΩäե֥ѥȽ񤭸ե֥ѥδ֤Ǥ f*ck ٺϸϰǤʤθ˴ؤξԤκϰ̣פ򤵤뤳Ȥˤʤ롥
Log-Likelihood Test ϰʾΤ褦˿ʤ뤬θԤʤˤäƤΤäƤɬפ롥̤ˤϡ׻٤ 5 򲼲륻뤬1ĤǤ⤢ˤϡ٤Ȥ롥 the Cochran rule ȸƤФƤ뤬꤭٤ʥ롼󵯤 Rayson, Berridge, and Francis (8) ˤС٤٤ͤͭտ 5% 13 1% 11 0.1% 8 Ȥͭտ 0.01% ꤹд 1 ˤѤ٤ΤǡRayson et al. ϥѥؤǴŪѤƤ3Ĥο˲äơ0.01% οб륫ͤ 15.13 ˤޤǤθ侩Ƥ롥
פˤϾܤʤɽ 2ʥ֡˥ѥ֤ǤӤȤǴñѤ뤳ȤǤ븡ȤơLog-Likelihood Test αϰϤϹ׻Τ Rayson Log-likelihood calculator ʤɤǤФ褤ܵϤΥڡεҤȥʸ򻲹ͤˤˡ
BNC Ѥ f*ck ϢʬۤθϡMcEnery et al. (264--86) Υǥ˾ܤ
ϢơϹԤʤʤäĤܥ֥ǰä gorgeous Ĵ ([2010-08-16-1], [2010-08-17-1],[2010-12-25-1]) ʤɤ⻲ȡ

Rayson, P., D. Berridge , and B. Francis. "Extending the Cochran Rule for the Comparison of Word Frequencies between Corpora." Le poids des mots: Proceedings of the 7th International Conference on Statistical Analysis of Textual Data (JADT 2004), Louvain-la-Neuve, Belgium, March 10-12, 2004. Ed. Purnelle G., Fairon C., and Dister A. Louvain: Presses universitaires de Louvain, 2004. 926--36. Available online at http://www.comp.lancs.ac.uk/computing/users/paul/publications/rbf04_jadt.pdf .
Rayson, P. and R. Garside. "Comparing Corpora Using Frequency Profiling". Proceedings of the Workshop on Comparing Corpora, Held in Conjunction with the 38th Annual Meeting of the Association for Computational Linguistics (ACL 2000), 1-8 October 2000, Hong Kong. 2000. 1--6. Available online at http://www.comp.lancs.ac.uk/computing/users/paul/phd/phd2003.pdf .
McEnery, Tony, Richard Xiao, and Yukio Tono. Corpus-Based Language Studies: An Advanced Resource Book. London: Routledge, 2006.

[ | ѥڡ ]

2011-03-12 Sat

#684. semantic prosody ʸˡƥ꡼ [semantic_prosody][grammar][corpus][intensifier]

äʿβϿ̤ˤĤޤơҼԤ˿ꤪ񤤿夲ޤ
[2011-03-03-1]εǡsemantic prosody ʸˡƥ꡼Ȥδ֤˴ϢȤǽ˸ڤϡhappenutterly ޤදո semantic prosody 򥳡ѥˤäĴ Partington ʸǻŦƤ뤳ȤǤ롥
Partington happen, set in, occur, come about, take place Ĵθ췲ˤ٤κϤ졤Τ unfavourable semantic prosody տ路ƤȤڵ󤲤ʺǤ unfavourable ʤΤ set in Ȥ(144) Ʊͤˡutterly, absolutely, perfectly, totally, completely, entirely, thoroughly Ĵ줾 semantic prosody 뤤 semantic preference Ф (148) ơĤθտ路Ƥ벻ˤϡfavourable vs. unfavourable ȤñʲʹΩǤϤʤ̤ˤʸˡƥ꡼ȤƸڤ褦ħΩͿƤȤȤ狼ä
Ū˸Сhappen non-factuality 򼨤ˡ䡤Ȥäʸˡƥ꡼ȤδͿǧ졤it is unclear why to see what ʤɤɽȤȤѤ뤳Ȥ¿ (140--41) ǡtake place Ϥष factuality 򼨤ͽꤵƤ뤳ȤºݤȤްդѤ뤳Ȥ¿ (143)
ոǤϡutterly unfavourable semantic prosody 򼨤ǤʤħԺߤѲɽ魯򽤾뷹 ( ex. utterly helpless / unable /forgotten / changed / different / destroyed ) Ʊϡtotally, completely, entirely ˤ⸫롥entirely ˤ (in)dependency Ȥƥ꡼ͿƤꡤentirely dependent / self-sufficient / isolated ʤɤѤ뤳Ȥ¿absolutely superlative ްդ򽤾 ( ex. absolutely delighted / splendid / appalling )
factuality, absence, change, dependence, superlative Ȥɤϡ̾ʸˡƥ꡼˴ϢƸڤ٥ΰ̣ä semantic prosody semantic preference ȤƸڤ̣ȿؤäƤ뤳Ȥ狼롥
ͤƤߤСäʸˡηӤĤȤϡʤʤ㤨СưϼȤǤѤʤȤѤ뤳Ȥ¿ʤɤȤ¤Τ褦˻ŦƤؽѼ˹ȿǤƤ롥ΰ̣ΰɽ魯줬³ that ư subjunctive ׵᤹ȤʸˡܤĹƤ ([2010-04-07-1]) äʸˡδطϱѸؤǤϤ褯ΤƤ¤ѥؤȤ٤Ʊ¤ˤɤ夤ȤȤѥؤι׸ϡfactuality absence ʤɤΥƥ꡼ 0 1 binary ȤƤǤϤʤprobabilistic ȤƼ갷ȤǤˤ褦˻פ롥
Ѹˤ뤤̻ؤδϡ줬ʸˡƥ꡼ȷӤĤǧˡġɤΤ褦ˤηӤĤΤ˶̣롥㤨 happen ϱѸˤΤĺ unfavourable non-factual ʴߤΤ⤷ˤΤ褦ʴߤӤӻϤ᤿ΤǤСΰ̣ξ ( semantic field ) ¾Ȥδط碌ƹͤɬפ롥ơȤδطȤȤˤʤСoccur ʤɼѸΰϤθʤФʤʤѸˤ̣ξκ semantic prosody ʸˡƥ꡼ؤηӤĤȤή줬ȤС⤷speculation ˤʤ㤨[2009-08-17-1]εǿ줿ȲˡߤȤδطˤή줬ʤ

Partington, A. "'Utterly content in each other's company': Semantic Prosody and Semantic Preference." International Journal of Corpus Linguistics 9.1 (2004): 131--56.

[ | ѥڡ ]

2011-03-05 Sat

#677. Ѹˤˡưο [auxiliary_verb][corpus][brown]

Ѹˡư ( modal auxiliary ) ηϤʣʤȤˤĤƤϡ[2010-07-22-1], [2010-01-20-1], [2009-07-01-1], [2009-06-25-1]εǿ줿ˡưϰưӤ־ο񤤤ðۤǤꡤ̣¿ͲƤΤDZѸˤ̤԰ʸǤäѸǤηŪʰƤ餺ʹȹͤ뤬ꤽ켫ΤʣǤ롥ѸˡưθϿ¿ηϤѲη򵭽ҤȤơThe Brown family of corpora ([2010-06-29-1]) Ѥ Leech et al. (Chapters 3--5) θ椬롥ä4 (pp. 71--90) Ǥϡפ11ˡư٤ѲܽҤƤ롥ʲϡ1961ǯƱѽ񤭸դɽ Brown LOB1991/92ǯƱѽ񤭸դɽ Frown F-LOB ˤꡤ30ǯ֤ˤ錄ˡư٤̻Ѳɽ路դǤLeech et al., p. 283 οɽȤ˺ˡͥǡϤΥڡHTML򻲾ȡwill ˤ 'll won't ʤɤξάޤࡥneed Ϲξޤࡥ

Frequencies of Modals in the Brown Family Corpora


ΤȤ30ǯδ֤ˡư٤äƤ뤳Ȥʬ롥٤θϡwould, will, may, should, must, shall, ought (to) p < 0.001 ˶ͭպ򼨤might, need(n't) p < 0.01 ζͭպ򼨤ϱξѼҤä᤿̤ѼʬĴȡAmE Τۤ BrE ⸺ٹ礤BrE η٤ɤƤ뤫Τ褦ʬۤ򼨤 (73)
̣ΤϡȤ٤㤤ˡưۤɸΨ礭 "bottom-weighting" (73) ηѻ뤳ȤΨʿѤ18.9%4ưǤߤ4.7%7ưǤߤ22.7%Ǥ롥äˡshall 43.5%ought (to) 37.5%need(n't) 31.6%ȤΨ
bottom-weighting طʤˤϡ"paradigmatic atrophy" ηŪಽ(80--81) ΤǤϤʤȻŦƤ롥ҤΤ褦ˡˡưϰư٤¿ԴǤ§ŪǤ롥;Τˤޤ礤Ƥꡤ¸ߤѲ⤭Ե§Ǥ롥shall 2;̾ȤƸ뤳ȤϤۤȤɤʤmayn't ȤϤƤޤǤ롥ˡư줬Ū "defective" ʸǤ뤳ȤͤСȤ櫓٤㤤ˡư줬äǽ˴٤ꡤޤޤ٤ˤʤäƤ椯ȤȤԻ׵ĤǤϤʤˡưκϥդ˸ۤñʸݤǤϤʤѥѤŪĴˤä礭Ĭή餫ˤ줿ȸ

Leech, Geoffrey, Marianne Hundt, Christian Mair, and Nicholas Smith. Change in Contemporary English: A Grammatical Study. Cambridge: CUP, 2009.

Referrer (Inside): [2015-04-22-1] [2014-12-02-1]

[ | ѥڡ ]

2011-03-04 Fri

#676. ѥθϤɤޤDzΩĤ [semantics][corpus][semantic_prosody][passive]

[2011-03-02-1], [2011-03-03-1]ε semantic prosody ꤢ붦ɽʼŪʡɾӤӤ븽ݤǤ롥semantic prosody ñʤΥ٥ˤȤɤޤ餺Ūʥ٥ˤ⸫롥㤨СStubbs (163--68) Ǥ be-passive Ф get-passive ΰ̣˴ؤ륳ѥѸ椬Ҳ𤵤Ƥꡤget Ѥư֤ϼ줬פȤʸ̮ʤ˾ˤäƤϼ줬פ˼ǤȤʸ̮ˤˤ˸Ȥ̤𤵤Ƥ롥
get-passive Ū semantic prosody ӤӤ䤹ȤȤϡ褫ʸˡǻŦƤȤѥĹ϶Ūʿ󶡤Ƥˤ롥Stubbs ĴǤϡbe-passive 25% "unpleasant" ʷ̤ްդ"pleasant" ްդΤ¿Ȥget-passive Ǥ60%ʾ夬 "unpleasant" ʷ̤ްդ"pleasant" ްդΤϤۤΤ鷺Ǥ롥̤ΥѥѤ̤θԤˤĴǤϡget-passage "unpleasant" ްΨäեѥ9ãȤ⤢ꡤget-passive Ū semantic prosody äƤ뤳Ȥ餫Ǥ롥Τ褦ʵҴŪʿͤˤ΢դcorpus semantics νפĹǤǤ롥
ѥˤä줿 get-passive ˴ؤ뤳θϡget-passive ޤŪʸβˤɤΤ餤ΩĤΤѥ줿Ȥʸͤ褦

I got praised for having a clean plate.


츫Ȥä "unpleasant" ްդϴޤޤƤʤget-passive ѤƤȤȤϡǤ "unpleasant" ްդᡤ餯Ūɤߤ׵ᤵƤȤȤʤΤѥˤθ뤳ȤϡŪ semantic prosody ȼäƤ get-passive ѤƤʾ塤⤤Ψ "unpleasant" ɤߤդ路"pleasant or neutral" ⳧̵ǤϤʤäΤ餳Ǥ㳰Ū "pleasant or neutral" ɤߤ⤷ʤפۤɤǤϾQŪΤäƤ뤳ȤȺʤѥθۤȤɳ褫ƤʤѥΥޤϡ̤㤫鷹õФȤդġβݾڤƤϤʤȤȤǤ롥ʸΤ˥ѥɽ̵ͭ٤Ĵ٤ȤȤŪ˹ԤʤäƤ뤬ǤĤפΤɽä顤٤äȤäơ줬ɬʸƳƤȤϸ¤ʤȤȤǤ롥ֻͤޤǤˡפǻߤޤäƤޤȤ¿äֻͤޤǤˡפǤϻͤˤʤʤȤ¿Τ
semantic prosody δȤ館ʤȡ붦ɽˤ semantic prosody δްդɤ٤ζ١괶ϤäƤС츫ȤΩŪŪʸ̮ʤɤŪʲӤӤȹͤΤ probability ͤȤƻФǤΤʤΤ
ġʸ̮ȽǤ٤ȸäƤޤФޤǤѥ̤ʸȤŪ˹׸ʤȤʤȡβͤ¤ƤޤΤǤϤʤStubbs ʸϡѥȲδطˤĤƾ嵭󵯤Ƥ뤬ˤĤƤ̵Ǥ롥

Stubbs, M. "Texts, Corpora, and Problems of Interpretation: A Response to Widdowson." Applied Linguistics 22.2 (2001): 149--72.

Referrer (Inside): [2011-03-20-1] [2011-03-11-1]

[ | ѥڡ ]

2011-03-03 Thu

#675. collocation, colligation, semantic preference, semantic prosody [semantics][corpus][collocation][semantic_prosody]

ε[2011-03-02-1]Ǽꤢ semantic prosody ˴Ϣꡥȸζطˤ4Ĥμब̤롥ʲMcEnery et al. (84--85, 149--52) 򻲾Ȥơ٤㤤Τ⤤Τؤ¤١줾γפ򵭤

(1) collocation: ùܤȸùܤȤδط
(2) colligation: ùܤʸˡƥ꡼Ȥδط
(3) semantic preference: ùܤȡ̣Ū˴Ϣ췲Ȥδط
(4) semantic prosody: Ụ߽̄Фùܤζط

(1) collocation ñ˸ȸ줬ȤطؤŪˤŪʳǰȹͤƤ롥ɤ٤٤äƶ "collocate" ƤȸʤȤǤΤ˴ؤơԤΤŪʴϰۤʤ see [2010-03-23-1], [2010-03-04-1] ) ̾ϡQŪˡֹ١פǤ collocation ȸƤǤ褦
(2) ̾ house ȺǤ٤Ƕ the a ʤɤδ줬뤬 collocation 򸦵椹Ǥޤ̣ͭǤʤ̾ǤдȶΤϼǤꡤhouse ˸ꤵ줿äǤϤʤcollocation ̣ͭʽѸȤݤĤˤϡhouse ȴΤ褦ʡʸˡƥ꡼δطɽ魯Ѹ줬ɬפȤʤ롥줬 colligation Ǥ롥
(3) semantic preference ϡ̣Ūͭ롤٤Ƕν˴ؤطǤ롥㤨Сlarge Ͽ̡Ϥɽ魯췲 ( ex. number(s), scale, part, quantities, amount(s) ) ȶutterly ħηǡ֤Ѳɽ魯췲 ( ex. helpless, useless, unable, forgotten; changed, different ) ȶ롥large utterly ϶ΰ̣ϰϤǤ롥
(4) semantic prosody Ϻε[2011-03-02-1]ǵ̤ǡ٤ɾȤäŪʰ̣߽ФطؤüԤΰռ˾ʤ줿ްդǤ뤳Ȥ¿semantic preference üʸȸ뤳ȤǤζܤɬΤǤϤʤ
μζǤ졤˴ؤܺ٤ʸŻҥѥǰ٤¿ʸ򽸤褦ˤʤäȤˤȯŸƤsemantic prosody θϡ̣ȯŸ˹׸뤳ȤϤޤǤʤ֤ζ̤餫ˤΤΩĤȤޤΤǸض伭ؤʬˤ׸뤳Ȥˤʤޤμθϸ̣ȶӤĤ븦ǤϤ뤬 utterly ȤδϢǼħηǡ֤ѲפȤ̣δͿͤȡpolarity modality Ȥäʸˡƥ꡼ȤδϢ⼨졤Ȥ⸫ơ֤뤳Ȥˤΰ̣夷Ƥ椯Ȥ˾ƤС̻Ūʸоݤˤʤ롥
semantic prosody ϡΤ褦˹ϤʱѤԤǤǤ롥McEnery et al. (84) ˺Ƕθν郎ΤǡͤޤǤ˰ʲƤ

Hunston, S. Corpora in Applied Linguistics. Cambridge: Cambridge UP, 2002.
Louw, B. "Irony in the Text or Insincerity in the Writer? The Diagnostic Potential of Semantic Prosodies." Text and Technology: In Honour of John Sinclair. Eds. M. Baker, G. Francis and E. Tognini-Bonelli. Amsterdam: John Benjamins, 1993. 157--76.
Louw, B. 2000. "Contextual Prosodic Theory: Bringing Semantic Prosodies to Life." Words in Context: A Tribute to John Sinclair on his Retirement. Eds. C. Heffer, H. Sauntson and G. Fox. Birmingham: U of Birmingham, 2000.
Partington, A. Patterns and Meanings. Amsterdam: John Benjamins, 1998.
Partington, A. "'Utterly content in each other's company': Semantic Prosody and Semantic Preference." International Journal of Corpus Linguistics 9.1 (2004): 131--56.
Schmitt, N. and R. Carter "Formulaic Sequences in Action: An Introduction." Formulaic Sequences. Ed. N. Schmitt. Amsterdam: John Benjamins, 2004. 1--22.
Stubbs, M. "Collocations and Semantic Profiles: On the Cause of the Trouble with Quantitative Methods." Function of Language 2.1 (1995): 1--33.
Stubbs, M. "Texts, Corpora, and Problems of Interpretation: A Response to Widdowson." Applied Linguistics 22.2 (2001): 149--72.


McEnery, Tony, Richard Xiao, and Yukio Tono. Corpus-Based Language Studies: An Advanced Resource Book. London: Routledge, 2006.

[ | ѥڡ ]

2011-03-02 Wed

#674. semantic prosody [semantics][corpus][collocation][semantic_prosody][terminology]

semantic prosody ϡǯΥѥؤζδˤä߽Ф줿ǰǤꡤȤƤܤ褦ˤʤäƤƱѥؤˤäܤ򽸤褦ˤʤä collocation Ȥ⿼ϢƤ롥Louw (57) ˤСsemantic prosody "a form of meaning which is established through the proximity of a consistent series of collocates" Ǥ롥⤦ʬ䤹Ȥ Crystal Ѥ褦

A term sometimes used in corpus-based lexicology to describe a word which typically co-occurs with other words that belong to a particular semantic set. For example, utterly co-occurs regularly with words of negative evaluation (e.g. utterly appalling). (428)


Ȥ utterly appalling 󤲤Ƥ褦ˡutterly ȤդϾˡŪɽ魯Ĵ롥¾ˡhappen set in ȤʶưԲʽɽ魯̾ȶ뤳Ȥ¿semantic prosody Ȥϡˤäƶ뤳Τ褦ʡְ̣βפΤȤؤμ礿뵡ǽüԤ٤ɾɽ魯ȤǤ롥¿Ūɾ˴ؤΤǤꡤŪɾϾʤʸԤȤƤϡŪʶ utterly ФƹŪʶ perfectly 󤲤褦ˡsemantic prosody collocation ȶӤĤƤ뤳ȤϡMcEnery et al. (83) ε󤲤Ƥ personal price 㤫餫Ǥ롥personal price ñȤǤϤɾΩŪ̾Ūʰ̣βȼ
ζˤä semantic prosody 줬ʬ夷Ƥȡζΰդ˰æ뤳Ȥˤä桼⥢ʤɤüʸ̤ɽ魯ȤǤ褦ˤʤ롥㤨СCobuild written corpus ˼Τ褦ʸ롥

Their relationship in fact was so complete that they were utterly content in each other's company.


semantic prosody ˴ؤ򤱤뤳ȤΤǤʤϡȸζˤäƤʤβʼŪʲˤΤ뤤Ū˳ƤΤȤǤ롥utterly ϤʤŪʲӤӤΤ䤤ФơŪʸȶ뤳Ȥ¿ä utterly ΤβӤӤ褦ˤʤäȤ뤫⤷ʤ⤽Ūʸȶ뤳Ȥ¿äΤϤʤʤΤ utterly ΤŪŪʲӤӤƤǤϤʤޤ˷ܤ褫褫˴٤äƤޤΤ褦ʾξȤơ(1) ŪŪ (2) ŪʸȤˤʶȤ2Ĥװߤ˺Ѥ̤ȤäȤⲺ򤫤⤷ʤŪǶᡤ -ish ŪʴްդγˤĤŪʸԤʤäˤȤäƤϡǺޤǤ롥McEnery et al. (84) ⤳˿Ƥ롥

It might be argued that the negative (or less frequently positive) prosody that belongs to an item is the result of the interplay between the item and its typical collocates. On the one hand, the item does not appear to have an affective meaning until it is in the context of its typical collocates. On the other hand, if a word has typical collocates with an affective meaning, it may take on that affective meaning even when used with atypical collocates. As the Chinese saying goes, 'he who stays near vermilion gets stained red, and he who stays near ink gets stained black' --- one takes on the colour of one's company --- the consequence of a word frequently keeping 'bad company' is that the use of the word alone may become enough to indicate something unfavourable . . . .


Crystal, David, ed. A Dictionary of Linguistics and Phonetics. 6th ed. Malden, MA: Blackwell, 2008. 295--96.
Louw, B. 2000. "Contextual Prosodic Theory: Bringing Semantic Prosodies to Life." Words in Context: A Tribute to John Sinclair on his Retirement. Eds. C. Heffer, H. Sauntson and G. Fox. Birmingham: U of Birmingham, 2000.
McEnery, Tony, Richard Xiao, and Yukio Tono. Corpus-Based Language Studies: An Advanced Resource Book. London: Routledge, 2006.

[ | ѥڡ ]

2011-02-23 Wed

#667. COCA 50ʻ̤γϡ [lexicology][corpus][french][loan_word][adjective][statistics][coca]

ε[2011-02-22-1]˰³COCA ( Corpus of Contemporary American English ) ˴Ťñ٥ꥹȤѤѥåȡǥϡǺǶˤʤäɲä줿50ΥꥹȤѤơƱͤʻ̳Ĵ٤ΥꥹȤϸФ ( lemma ) ˴Ť5000졤ΥꥹȤϸ ( word form ) ˴Ť50Τˤ497187ˤǡʤۤʤ뤳Ȥդ
ȤۤƱȤ2줺ĤdzڤꡤL1L25ޤǤγΤ줾ˤ noun, verb, adj., adv., others 5ʬʻ̳ФʿͥǡϤΥڡHTML򻲾ȡ

Form-Based POS Ratios by COCA

L612٥դ꤫ʻΨϰȤäƤ褤L1734٥դ꤫ưϤޤΤˤʤ뤬礭Ƥߤȡʤʤ餷ƤߤȡľΥ٥뤫礭æƤʤ
[2011-02-16-1]ε衤ƻΨˤʤäƤ뤬ΥǡΤ׻ȡ0.1738ȤͤϤ줿 lemma ĴǤ0.1678ä顤ͤ˶Ƥ롥̾ư lemma word form Ψϡ̾줬 0.5086 : 0.6985ư줬 0.2000 : 0.1065 礭ۤʤΤǡƻ 0.1678 : 0.1738 Ȥ϶⤷ʤlemma word form ʻ̳ˤϰۤʤ뷹Τ⤷ʤǤ絬ϤĴ٤ȰȸƤӤ֤и뤳ȤϳΤʤ褦
[2011-02-16-1]εǿ줿褦ˡѸΥե󥹼ѸˤƻΨ0.1768ä0.1738ȹƤ뤬ޤǰ㤦Τǡľܤδط뤳Ȥ̵Ǥ롥ȤȺĴϡ[2011-02-16-1]ĴȤ̵ط˻Ϥ᤿ΤǤ롥Ȼפ뤳η̤ϡŪǤϤ롥ѸäȤ̾줬ŪʤϤͽۤƤΤΡե󥹸ťΥɸ줫Ϥ褽γηƻʤ줾 lemma Ĵ0.17680.1817ˤѤƤơΨϻ夬ۤʤȤϤѸΨȶƤ롥ѸΤˤΨȼѸäˤΨƤȤȤϡ⤷ǤʤȤ顤̣Τե󥹼ѸäťΥɼѸäѸŬ褦ʼΨDZѸäϤȤȤϡΥѥåȡǥη̤Ƥΰݤ˴Ť speculation ˤʤʻ̳ȤܤƤ

[ | ѥڡ ]

2011-02-22 Tue

#666. COCA 5000ʻ̤γϡ [lexicology][corpus][statistics][n-gram][coca]

COCA ( Corpus of Contemporary American English ) ˴ŤƼåꥹȤ Corpus-based word frequency lists, collocates, and n-grams Ǥ롥ΤʤǺǤŪʥꥹȤκ5000ꥹǤ롥󤵤ƤΤϸФ ( lemma ) ñ̤ǡ̤ϥѥ˸٤ʬδؿǷ׻Ƥ롥UCREL CLAWS7 Tagset ʻ쥳ɽ˴ŤƤʻͿƤꡤʻ̤٤ʤɤڤʬϤ뤳ȤǤ롥
ϡ500줴Ȥ˶ڤä٤ι⤤L1L10ޤǤγߤ줾γˤʻ̳Фʻϳ ( open class ) 濴Ȥnoun, verb, adj., adv., others 5ʬȤʿͥǡϤΥڡHTML򻲾ȡ

Lemma-Based POS Ratios by COCA

1ɤγǤ̾줬ȾƤΤͽۤǤȤ2ʹߤ̾γ礬פäۤɿӤƤʤȤʬäưȷƻ줬ȾγǤ⤪褽γ³ƤΤͽ۳äΤȤơ5000ꥹȤ˸¤С̾줬ȴĤĤ⡤ΨϤ褽ݤƤȤ褦͡ưƻƤߤ褽Τ500ʹߤȸƤ褵
[2011-02-16-1]εѸΥե󥹼Ѹʻ̳ߤΤȤƤηƻΨ0.1768äθѸκ5000ǤϡΤȤƤηƻΨ0.1678٤ựΤͤɤʬʤѸʸ졩ˤˤʻΨΡְ괶פΤ褦ʤΤϤΤ
COCA ˴Ťΰʳ˥饤ǤѱñꥹȤˤĤƤ[2010-03-01-1]ε򻲾ȡɽѤ̤ΥѥåȡǥȤƤϡñβ򰷤ä[2010-04-17-1]ε򻲾ȡ

[ | ѥڡ ]

2011-01-29 Sat

#642. OED ΰѥǡ򥳡ѥȤƻȤ뤫 (4) [oed][corpus][statistics]

[2010-10-15-1]ε˴ϢơBrewer ʸ­ε OED ΰѿ̤˥ղΤǤä˸ä򼨤ƤսǼǤʲ˼

OED Quotations per Decade by Brewer (Marked)

Brewer (58) ˤȡ(1)--(5) γä OED ԻװˤȤ礭Ȥ롥줾λϰʲ̤Ǥ롥

(1) 1291--1300ǯá1470ǯˤĤƤϤФХƥȤǯ夬ǤꡤΤ褦ʾˤصξüǯꤹȤԽˤäޤäˤλˤĤƤϡRobert of Gloucester (1297ǯ3222) Cursor Mundi (1300ǯ10771 OED ˤѿ2̤κ) 顤ʤ꽸Ū˰ѤޤƤȤ⤢롥
(2) 1391--1400ǯá(1) ƱͤȤͳ˲äTrevisa (1387/98ǯ6750) ̤˼ޤƤȤ𤬤롥
(3) 1521--1530ǯáPalsgrave Lesclarcissement (1530ǯ5418) ̤ΰѤˤꡤȾ롥
(4) 1581--1600ǯáShakespeare (33304) αƶ礭
(5) 1631--1660ǯá餯̿ΥѥեåȤ¿ΰѤƶƤ롥

5äˤĤƤǤԽطʤŪΤäƤȡOED ΰѥǡλȤʾʤȤ⤽λˤѤäƤȻפ⤷补Ϣ뵭ȤƤϰʲ򻲾ȡ

[2010-10-10-1]: #531. OED ΰѥǡ򥳡ѥȤƻȤ뤫
[2010-10-14-1]: #535. OED ΰѥǡ򥳡ѥȤƻȤ뤫 (2)
[2010-10-15-1]: #536. OED ΰѥǡ򥳡ѥȤƻȤ뤫 (3)

Brewer, Charlotte. "OED Sources." Lexicography and the OED: Pioneers in the Untrodden Forest. Ed. Lynda Mugglestone. Oxford: OUP, 2000. 40--58.

Referrer (Inside): [2020-09-29-1] [2015-03-29-1]

[ | ѥڡ ]

2011-01-14 Fri

#627. 2Ѽ֤̻ӤˤäŪۤ෿ [language_change][speed_of_change][corpus][brown][ame_bre]

[2010-06-29-1]εǤߤ褦ˡThe Brown family of corpora 4ѥ ( Brown, Frown, LOB, F-LOB ) Ѥ뤳ȤˤäƱѸαѼ֤30ǯ֤ۤɤ̻Ѳ٤뤳ȤǤ롥Τ褦˿ꤹ­Ӳǽ򼨤ʣΥѥѤ̻ "diachronic comparative corpus linguistics" (Leech et al. 24) ȸƤФƤꡤߤ30ǯۤɤδֳ֤򤢤ѼΥѥ̤ξظäԻƤ椯ΤȻפ롥
ϰѼǯȤ2ĤΥѥ᡼ˤäܤ٤κˤĤơŪʲʣꤦ롥Brown family ξˤϤɤΤ褦ʲ᤬뤫Mair (109--12) Ƥ2Ѽ֤̻ӤˤäŪۡʤ̵ͭˤ෿ ( "typology of contrasts" ) Ѥǰʲ˼"=" Ѳνȯ"+/-" ѲȤ򼨤

(1) nothing happening
BrE: = =
AmE: = =

(2) stable regional contrast
BrE: = =
AmE: +/- +/-

(3) parallel diachronic development
BrE: = +/-
AmE: = +/-

(4) convergence: Americanization
BrE: +/- =
AmE: = =

(5) convergence: 'Britishization'
BrE: = =
AmE: +/- =

(6) incipient divergence: British English innovating
BrE: = +/-
AmE: = =

(7) incipient divergence: American English innovating
BrE: = =
AmE: = +/-

(8) random fluctuation
BrE: = +/-
AmE: +/- +/-

(1), (8) ϺǤ¿ѻԤδؿʤʿޤʥפκۡʤηǡˤǤ롥(2) ϳΩ줿ưαƺ㤨 <honour> vs. <honor> ֻ got vs. gotten λѤȤʤ롥(3) Mair Ǥϵ󤲤Ƥʤ(4) Americanization λ㡤㤨 help 褦ˤʤäƤƤ뷹פ⤫٤뤳ȤǤʤ BrE ǤΤηϤ٤Ƥ Americanization ˵Ȥ櫓ǤϤʤˡ(5) ˤޤ 'Britishization' Ǥ롥㤨 AmE Ǥνưɽ have got to ι BrE ˸ƤǽȵƤ롥(6) ϡBrE prevent "O + from + V-ing" ǤϤʤ "O + V-ing" 򹥤򤹤褦ˤʤФƤ뷹˵󤲤롥(7) ϡAmE begin to Ǥʤ V-ing ٤ޤФƤ뷹Ȥʤ롥
ŪˤϡѲ®٤θʤФʤʤ㤨 (3) Τ褦ξѼƱ̻ѲƤǤ⡤Ѽ֤Ѳ®٤˺з̤ȤʿԤˤϤʤʤ嵭෿®٤Ȥȡ˺٤ʬɬפˤʤϤǤ롥Τ褦ʣʲϻĤäƤ뤬2Ѽ2Ӥ "diachronic comparative corpus linguistics" ŪȤơ嵭 "typology of contrasts" ͭѤ󡤤ΥݥϡBrE AmE ˤ30ǯۤɤȤû֤̻ѲǤʤʹߤξѼ̻Ūȯã򵭽ҤǥȤƤͭǤ롥ϡ[2010-10-09-1]εǰäѸ convergence divergence ˤŬѤǤȻפ롥

Leech, Geoffrey, Marianne Hundt, Christian Mair, and Nicholas Smith. Change in Contemporary English: A Grammatical Study. Cambridge: CUP, 2009.
Mair, Christian. Three Changing Patterns of Verb Complementation in Late Modern English: A Real-Time Study Based on Matching Text Corpora." English Language and Linguistics'' 6 (2002): 105--31.

[ | ѥڡ ]

2010-12-25 Sat

#607. Google Books Ngram Viewer [corpus][web_service][ame_bre][google_books][n-gram][statistics][frequency][lexicology]

Google Τѥġ󶡤ƤGoogle Books Ngram Viewer Google Labs εϤȲǽ礭˶ä2004ǯ1500ܤǥ벽Ƥ Google Υ֥åȤȤʤ520ܡ5000򥳡ѥѸΤۤե󥹸졤ɥĸ졤졤ڥ졤줬ޤޤƤ뤬ѸǤ British English, American English, English, English Fiction, English One Million 饵֥ѥǤ롥ħϡꤷ5ޤǤθ٤51500--2008ǯˤˤ錄äפդɽƤ뤳ȤGoogle θεˤ롥
Ϥ礭ƥѥȤƤɤɾ٤ʬʤҤȤޤϤdzڤ嵭εˤĤΥץ뤬뤬ѸŪʴؿץȤ burnt burned ʬӤäΤǡEnglish, American English, British English 3֥ѥ򥰥դФƤߤ
ˡǯ٤´ΰäҼڤ̤ AmE on the street, BrE in the street ȤֻѤκۤ Google Books Ngram Viewer dzǧƤߤAmerican English British English Τ줾Υ֥ѥϤ줿դϰʲ̤ꡥ

in the street and on the street by Google Books Ngram Viewer

in on ϶ΰ̣ʡֳϩǡפּȤơפˤʤɤˤ¸뤿ñʷ֤ӤǤԽʬϤĤ롥
[2010-08-16-1], [2010-08-17-1]εǰä gorgeous ˤĤƤĴ٤Ƥߤ19ˤήԤäƤ20ˤܤǤäηƻ줬American English ˤ1980ǯʹߡƤ֤ƤƤ褯狼롥British English ǤĴ
ѥذ̤ˤ뤬ġλѤϥǥǤ롥ʸŪʴϡ[2009-12-28-1]εǾҲ𤷤 American Dialect Society ˤ "Words of the Century" "Words of the Millennium" ΥΥߥ͡ȸ򸡺ƤߤȤ⤷
¾Υ饤󥳡ѥˤĤƤ[2010-11-16-1]򻲾ȡ

[ | ѥڡ ]

2010-11-16 Tue

#568. ѥȱѸ쥳ѥ [corpus][link][representativeness]

츦ˤ corpus ֥ѥפ͡Ƥ뤬McEnery et al. ʷǤ롥

. . . a corpus is a collection of (1) machine-readable (2) authentic texts (including transcripts of spoken data) which is (3) sampled to be (4) representative of a particular language or language variety.


(1) (2) ˤĤƤϤ褽Դ֤˥󥻥󥵥뤬(3) (4) ˤĤƤϲä "sampled" 뤤 "representative" ȤߤʤˤĤ͡ʰո롥ڤˤƤ뤳ȤǤ
ڤ˱Ѹ쥳ѥˤϡ饤ΤΤǤ롥ʲϡϿɬפʤΤ⤢뤬˥饤ǴؤѤǤѸ쥳ѥ

British National Corpus ʤĤΥ󥿡ե󶡤Ƥ

* BNC ( The British National Corpus )
* BNCweb ̵Ͽ
* BYU-BNC ̵Ͽ

BYU Corpora Brigham Young University, Mark Davies 󶡤Τ¾Υ饤󥳡ѥ

* COCA ( Corpus of Contemporary American English ) ̵Ͽ
* COHA ( Corpus of Historical American English ) ̵Ͽ
* TIME Magazine Corpus of American English ̵Ͽ

Cobuild Concordance and Collocations Sampler

¾ܥ֥Ǥϥѥطε򤤤ȷǺܤƤΤǡͤˤ줿

hellog Υѥν󵭻: [2010-09-15-1]
hellog ΥѥϢ: corpus
hellog BNC Ϣ: bnc

McEnery, Tony, Richard Xiao, and Yukio Tono. Corpus-Based Language Studies: An Advanced Resource Book. London: Routledge, 2006.

[ | ѥڡ ]

2010-10-15 Fri

#536. OED ΰѥǡ򥳡ѥȤƻȤ뤫 (3) [oed][corpus][statistics]

[2010-10-10-1], [2010-10-14-1]˰³OED ΰѥǡꡥϡä˺ε[2010-10-14-1] (2), (3) Ǽ夲ǯ̰ѿ⤭ߤռǡͤ򥰥դ˻вƤȹͤ
Brewer 10ǯȤ OED ΰѿοܤĴ٤Ƥꡤºݤ˥ղ⤷Ƥ (48--49) ʸ󼨤Ƥ륰դ1470ǯ򶭤ʬƤꡤ٤ߤ˰ۤʤäƤΤӤˤϤؤǤ롥ǡʲ٤·դƺƤߤBrewer ˤϥպΤȤˤʤͥǡͿƤʤΤǡդܸƤǿͤɤ߽Ф˺ʢ ϼ OED DzƿФФΤɡˡäơ˼ƤΤϤޤǷȤ館뤿ΤΤȤƻͤޤǤˡ

OED Quotations per Decade by Brewer

OED ̻ѥȤѤˤϡä˰ѿϤŪ㤫ä⤫äꤹΰѤݤդɬפǤ롥ΥդϡκݤΤȤƻȤ줿

Brewer, Charlotte. "OED Sources." Lexicography and the OED: Pioneers in the Untrodden Forest. Ed. Lynda Mugglestone. Oxford: OUP, 2000. 40--58.

Referrer (Inside): [2015-03-29-1] [2011-01-29-1]

[ | ѥڡ ]

2010-10-14 Thu

#535. OED ΰѥǡ򥳡ѥȤƻȤ뤫 (2) [oed][corpus]

[2010-10-10-1]εǤϡHoffmann ʸ򻲾ȤơOED ΰѥǡϼ㴳դɬפʬ˥ѥȤʤꤦΤǤϤʤȤ򸫤ǡOED ΰѤϼ㴳ǤϤʤդʧʤȴʤȤ⤬롥Brewer ˤСOED ΰѥǡ򡤳ƻɽ륳ѥȤߤʤȤˤϿŤǤ٤ȤBrewer ʸ򻲾ȤĤ͡ʾڵ󤲤ƵƤ뤬ʤΤ򲼤ˤޤȤƤߤ롥

(1) ʸغȡʸغʤΰѤ礤¿ѿȥå5κȤϡShakespeare, Walter Scott, Milton, Wycliffe, ChaucerShakespeare ΥСΨ100%˶ᤤȸ졤ѿ33304롥5̤ Chaucer ΰѤ11902㡥ѿȥåפκʤϡͽ̤2̤1300ǯ˽񤫤줿Ĺ Cursor Mundi 12772롥ͭ̾ʺȡʤˤĤƤϥ󥳡󥹤䤹ˡѤѤ䤹Ȥ𤬤Ȥ (45--47) ѤϸɽƤȤ⡤ԻԤɽ路ƤȤ٤Ǥ롥

Any inferences drawn from the OED coverage about the significance of these writers for the development and illustration of the English lexicon are flawed ones: the exceptionally full representation of their language in the dictionary is due at least as much to the lexicographers' consultation of the concordances as to the intrinsic qualities of these writers' diction. (51)


(2) ѿǯ̤˥ץåȤ c1581--1610 ˰Ѥ޷Ƥ롥ޤ19ȾѤʤФƤ롥ˤĤƤ[2010-10-10-1] (4) Ǥ⿨줿ԤλˤĤƤ Shakespeare ΰѤ¿ȤȿϢƤꡤɬ⤽λθɽƤȤȤˤϤʤʤΤǤϤʤ (47, 58) ԤλˤĤƤϡOED ΤλǤꡤɬŪưפ˼ŵο¿Ǥ롥

(3) 15Ǥ 1291--1300, 1391--1400 λ˰ѤΥԡ뤬1Ĥˤǯ夬ΤʺʤˤĤƤ϶ڤΤ褤Ѥܤڤ夲ڤ겼ꤹ뤳Ȥꡤ줬ȿǤ줿̤Ȥ̤ͳȤƤϡ1300ǯ Robert of Gloucester 3222ˤ Cursor Mundi 10771ˤ1400ǯ Trevisa 6750ˤ椷Ǥ (57--58)

(4) OED ˺Ѥ븫ФϱѸΥܥƥɼԤˤñȤΥ⤬ˤʤäƤ뤬ܥƥ̤Ǥʤ̤Ǥʤ̣äդƽ褦˻ؼƤ". . . this resulted in partial reading and uneven representation of sources" (50).

(5) OED ˤϽѸμľܰѤƤ븫Ф줬¿뤬μθФ줬٤ƼϿƤ櫓ǤʤФ줬򤵤Ƥפ롥Ĵˤȡ1/5ۤɤ OED ˤϼϿ줺ڤΤƤ줿ȤǤϡԻԤŪȽǡ餯19οʲѤ΢Ǥ줿ϼŪȽǤäƤȹͤ (52--52)

[2010-10-10-1]Ȥ碌 OED ΰѥǡ򥳡ѥȤƤߤʤƤ褤ɤˤĤƻξ򸫤1000ǯ˱Ѹ򥫥С밷䤹̻ѥ¾˸Ƥʤʾ塤˵󤲤褦ռ OED դѤ롤ȤȰʳϤʤ褦˻פ롥

Brewer, Charlette. "OED Sources." Lexicography and the OED: Pioneers in the Untrodden Forest. Ed. Lynda Mugglestone. Oxford: OUP, 2000. 40--58.

[ | ѥڡ ]

Powered by WinChalow1.0rc4 based on chalow