상세 보기
초록
This paper aims to propose that Benford's Law, non-uniform distribution of the leading digits in lists of numbers from many real-life sources, also appears in linguistic texts. The rst digits in the frequency lists of morphemes from Sejong Morphologically Analyzed Corpora represent non-uniform distribution following Benford's Law, but showing complexity of numerical sources from complex systems like earthquakes. Benford's Law in texts is a principle re ecting regular distribution of low-frequency linguistic types, called LNRE(large number of rare events), and governing texts, corpora, or sample texts relatively independent of text sizes and the number of types. Although texts share a similar distribution pattern by Benford's Law, we can investigate non-uniform distribution slightly varied from text to text that provides useful applications to evaluate randomness of texts distribution focused on low-frequency types.
키워드
- 제목
- 언어 텍스트에 나타나는 벤포드 법칙: 원리와 응용
- 제목 (타언어)
- Benford's Law in Linguistic Texts: Its Princi- ple and Applications
- 저자
- 홍정하
- 발행일
- 2010
- 저널명
- 언어와 정보
- 권
- 14
- 호
- 1
- 페이지
- 145 ~ 163