Empirical verification of the probabilistic concept of the linguistic norm (a preliminary study based on Polish material)

Michał Szczyszek

Uniwersytet im. Adama Mickiewicza w Poznaniu
https://orcid.org/0000-0002-0253-7296


Abstract

The article presents empirical verification of a methodological proposal within the field
of normativity studies, i.e. verification of a probabilistic concept of the linguistic norm.
This framework posits that linguistic structures employed in texts may exhibit varying
degrees of probabilistic alignment with a Normalized Language Pattern. This pattern is conceptualized as a statistical centre derived from Gaussian models – specifically,
as the ‘area’ of a normal distribution oscillating around the median and spanning one
or two percentiles; the area of normalized language use would fall within the boundaries
of the first standard deviation (sigma). Following the procedure outlined in the theoretical
framework, analyses were conducted on four linguistic constructions traditionally classified
as errors in the Polish language: *przyszłem/ przyszedłem [I came]; kartka papieru [a sheet
of paper]; *cofać się z powrotem [to go back]; *włanczać/ włączać [to switch on]. Results
of these analyses appear to tentatively confirm the operability of the proposed probabilistic
model. Furthermore, they underscore the necessity of developing a digital linguistic tool
– a ‘normaliser’ – utilizing Gaussian formulas and statistic instruments to calculate the
normal distribution of linguistic phenomena. Such a tool would automate the process
of codifying linguistic phenomena in accordance with the probabilistic concept of the
linguistic norm.


Keywords:

linguistic norm, probabilistic concept of the linguistic norm, erification of the probabilistic concept of linguistic norms, normaliser


```Bąba S. (1981): Z zagadnień współczesnej normy językowej. „Studia Polonistyczne” IX, s. 27–36.   Google Scholar

Bedyńska S., Cypryańska M. (red.) (2013): Statystyczny drogowskaz. 1: Praktyczne wprowadzenie do wnioskowania statystycznego. 2: Praktyczne wprowadzenie do analizy wariancji. Warszawa.   Google Scholar

Buttler D., Kurkowska H., Satkiewicz H. (1971): Kultura języka polskiego. Zagadnienia poprawności gramatycznej. Warszawa.   Google Scholar

Coșeriu E. (1952): Sistema, norma y habla. Montevideo.   Google Scholar

Durkheim É. (2000): Zasady metody socjologicznej. Tłum. J. Szacki. Warszawa.   Google Scholar

Shapiro S.S., Wilk M.B. (1965): An Analysis of Variance Test for Normality (Complete Samples). „Biometrika” Vol. 52, nr 3/4, s. 591–611.
Crossref   Google Scholar

Szczyszek M. (2025a): Akwizycja normy językowej. Czy zachodzi, jak i kiedy? Na przykładzie systemu słowotwórczego i – częściowo – leksykalnego. „Język Polski” CV/2, s. 30–40.
Crossref   Google Scholar

Szczyszek M. (2025b): Jak rozumieć trzeba normę językową? Kodyfikacja – od powtarzalności metanormotwórczej ku innowacyjnym możliwościom korpusologicznym i AI: probabilistyczna koncepcja normy językowej. [W:] Kultura komunikacji językowej – powtarzalność i innowacyjność. Tom dedykowany pamięci Profesora Stanisława Bąby w 10. rocznicę śmierci. Red. A. Piotrowicz-Krenc, M. Witaszek-Samborska, K. Skibski. Poznań, s. 267–281.   Google Scholar

Wawrzynek J. (2007): Metody opisu i wnioskowania statystycznego. Wrocław.   Google Scholar

Zabrocki L. (1963): Wspólnoty komunikatywne w genezie i rozwoju języka niemieckiego. Cz. I: Prehistoria języka niemieckiego. Wrocław–Warszawa–Kraków.   Google Scholar

Zimny A. (2010): Statystyka opisowa. Materiały pomocnicze do ćwiczeń. Konin.   Google Scholar


Published
2026-06-30

Cited by

Szczyszek, M. . (2026). Empirical verification of the probabilistic concept of the linguistic norm (a preliminary study based on Polish material). Papers in Linguistics, 28(2), 159–177. https://doi.org/10.31648/pj.12516

Michał Szczyszek 
Uniwersytet im. Adama Mickiewicza w Poznaniu
https://orcid.org/0000-0002-0253-7296