Search | Preprints.org

Working Paper ARTICLE

DASTEX: a New Readability Formula based on Semantic Complexity of Text

Mohammad Reza Besharati, Mohammad Izadi

Subject: Computer Science And Mathematics, Algebra And Number Theory Keywords: Semantic Complexity; Semantics; Text Complexity; Readability Formulae

Online: 6 September 2021 (13:33:34 CEST)

Show abstract| Download PDF| Share

Preprint ARTICLE | doi:10.20944/preprints202312.0900.v1

A Mathematical Structure Underlying Sentences and Its Connection with Short–Term Memory

Emilio Matricciani

Subject: Computer Science And Mathematics, Other Keywords: Alphabetical Languages; Extended Short–Term Memory; Human Communication; Human Mind; Sentences: Mathematical Modeling; Universal Readability Index

Online: 12 December 2023 (15:36:55 CET)

Show abstract| Download PDF| Share

Preprint ARTICLE | doi:10.20944/preprints201811.0505.v1

Italian Throughout Seven Centuries of Literature: Deep Language Statistics And Their Relationship With Miller’s 7∓2 Law and Short−Term Memory

Emilio Matricciani

Subject: Social Sciences, Language And Linguistics Keywords: Italian, readability, GULPEASE, literature, statistics, characters, words, sentences, punctuation marks, short−term memory, word interval, time interval

Online: 20 November 2018 (15:32:11 CET)

Show abstract| Download PDF| Share

Statistics of languages are calculated by counting characters, words, sentences, word rankings. Some of these random variables are also the main “ingredients” of classical readability formulae. Revisiting the readability formula of Italian, known as GULPEASE, shows that of the two terms that determine the readability index G – the semantic index G_C, proportional to the number of characters per word, and the syntactic index G_F, proportional to the reciprocal of the number of words per sentence −, G_F is dominant because G_C is, in practice, constant for any author throughout seven centuries of Italian Literature. Each author can modulate the length of sentences more freely than he can do with the length of words, and in different ways from author to author. For any author, any couple of text variables can be modelled by a linear relationship y=mx, but with different slope m from author to author, except for the relationship between characters and words, which is unique for all. The most important relationship found in the paper is, in author’s opinion, that between the short−term memory capacity, described by Miller’s “7∓2 law”, and the word interval, a new random variable defined as the average number of words between two successive punctuation marks. The word interval can be converted into a time interval through the average reading speed. The word interval is spread in the same of Miller’s law, and the time interval is spread in the same range of short−term memory response times. The connection between the word interval (and time interval) and short−term memory appears, at least empirically, justified and natural, and should further investigated. Technical and scientific writings (papers, essays etc.) ask more to their readers. A preliminary investigation of these texts shows clear differences: words are on the average longer, the readability index G is lower, word and time intervals are longer. Future work done on ancient languages, such as Greek or Latin, could bring us a flavor of the short term−memory features of these ancient readers.

Preprint ARTICLE | doi:10.20944/preprints202310.1661.v1

Is Short-Term Memory Made of Two Processing Units? Clues from Italian and English Literatures Down Several Centuries

Emilio Matricciani

Subject: Computer Science And Mathematics, Information Systems Keywords: Alphabetical Texts; Human Communication; Human Mind; Information; Linguistic Communication Channels; Miller’s Law; Processing; Sentence Modeling; Short-Term Memory; Universal Readability Index

Online: 25 October 2023 (16:17:13 CEST)

Show abstract| Download PDF| Share

Preprint ARTICLE | doi:10.20944/preprints201811.0149.v1

A Mathematical Analysis of Maria Valtorta’s Mystical Writings

Emilio Matricciani, Liberato De Caro

Subject: Arts And Humanities, Religious Studies Keywords: Confidence tests, dictations, Jesus Christ, Maria Valtorta, mystics, punctuation marks, readability index, sentences, semantic index, syntactic index, text characters, Virgin Mary, visions, words, word interval.

Online: 7 November 2018 (09:06:01 CET)

Show abstract| Download PDF| Share

Search Results

5 articles found