trustme.bro/r/…
✓ checked
trust me, bro:
here is the receipt.
the claim
Word frequencies exhibit universal cross-lingual patterns
the verdict
INSUFFICIENT LEANING
refutedsupported
the weight of evidence
3 sources for · 0 against

Available sources touch on general cross-lingual patterns and word embeddings, providing partial context for universal linguistic regularities without directly settling specific universal patterns in word frequencies.

Evidence for · 3
2020 · cited by 4
Grammatical gender is assigned to nouns differently in different languages. Are all factors that influence gender assignment idiosyncratic to languages or are there any that are universal? Using cross-lingual aligned word embeddings, we perform two experiments to address these questions about language typology and human cognition. In both experiments, we predict the gender of nouns in language X using a classifier trained on the nouns of language Y, and take the classifier’s accuracy as a measure of transferability of gender systems. First, we show that for 22 Indo-European languages the transferability decreases as the phylogenetic distance increases. This correlation supports the claim that some gender assignment factors are idiosyncratic, and as the languages diverge, the proportion of shared inherited idiosyncrasies diminishes. Second, we show that when the classifier is trained on two Afro-Asiatic languages and tested on the same 22 Indo-European languages (or vice versa), its performance is still significantly above the chance baseline, thus showing that universal factors exist and, moreover, can be captured by word embeddings. When the classifier is tested across families and on inanimate nouns only, the performance is still above baseline, indicating that the universal factors are not limited to biological sex. c⃝2020 Association for Computational Linguistics https://doi.org/10.18653/v1/P17 265 Cross-lingual Embeddings Reveal Universal and Lineage-Specific Patterns in Grammatical Gender Assignment Hartger Veeman Uppsala University Department of linguistics and philology Box 635, 75126 Uppsala hartger.veeman.7544@student.uu.se Aleksandrs Berdicevskis University of Gothenburg Spr˚akbanken Box 200, 40530 Gothenburg aleksandrs.berdicevskis@gu.se Marc Allassonni`ere-Tang University Lyon 2 Lab Dynamics of Language 14 avenue Berthelot 69363 Lyon marc.tang@univ-lyon2.fr Ali Basirat Uppsala University Department of linguistics and philology Box 635, 75126 Uppsala ali.basirat@lingfil.uu.se Abstract Grammatical gender is assigned to nouns dif- ferently in different languages. Are all fac- tors that influence gender assignment idiosyn- cratic to languages or are there any that are uni- versal? Using cross-lingual aligned word em- beddings, we perform two experiments to ad- dress these questions about language typology and human cognition. In both experiments, we predict the gender of nouns in language X using a classifier trained on the nouns of lan- guage Y , and take the classifier’s accuracy as a measure of transferability of gender systems. First, we show that for 22 Indo-European lan- guages the transferability decreases as the phy- logenetic distance increases. The classifier’s ability to classify the test nouns is interpreted as an indication of the transferability of grammatical gen- der system from the source language to the target language (i.e., the higher the accuracy is, the more transferable the gender systems are). The entire setting is founded on the cross-lingual represen- tation of words, providing for knowledge transfer between the gender classification models across languages. The embeddings are used to represent nouns in both source and target languages. When the classifier is applied to related languages, does its success depend on how related they are (if yes, this in an indication that some factors are not uni- versal)? Does gender transfer work in the same way for all nouns or are there differences between certain noun classes? 2 Experimental materials and settings In this section, we present the languages involved in this study along with the source of our data. Then, we provide an overview of the cross-lingual word embedding method and the settings of the classifier used for gender transfer. 2.1 Materials Two sources of data are selected for each language. Second, word embeddings are se- lected from pre-trained cross-lingual embeddings published on the fastText website (Joulin et al., 2018).2 Further details about the embeddings are provided in the following subsection. We selected all languages that have grammatical gender and are present in both data sources, with the exception of Albanian due to its small treebank size and Norwegian because in pilot experiments, our classifier showed unexpectedly poor perfor- mance for reasons we were not able to establish. This results in the selection of 24 languages that are shown in Table 1. 273 5 Conclusion This study investigates how grammatical gender is transferable across languages from a transfer learn- ing point of view. The cross-lingual word embed- dings are considered as the source of knowledge shared between languages from which the gram- matical gender of nouns are predicted using a multi- layer perceptron. The empirical results reveals that there exist some universal and lineage-specific pat- terns in the grammatical gender assignment. First, our analysis of gender transfer between Afro-Asiatic and Indo-European languages indi- cated that partly successful gender transfer is pos- sible between non-related languages. While we address the universality of gender assignment cross-linguistically, our data is restricted to languages from two families and our word embeddings are trained on data from specific domains. Additional data from a more diverse sam- ple is needed to further confirm our observations. Furthermore, we Linguistically, grammatical gender is strongly tied to the semantic and for- mal properties of nouns. Since the cross-lingual word embeddings used in this study encode both the formal and semantic information, we cannot disentangle the relative contributions of form and semantics to gender transfer. Finally, it should be mentioned that an important line of research in modern NLP focuses on gender bias present in naturally occurring texts (Caliskan et al., 2017; Gonen et al., 2019). The combina- tion of these questions and approaches with our perspective might become an interesting research direction.
See more details
The analysis

rails:sufficiency:partial_only:for=0+3p:against=0+0p | v55:multi_partial_one_side:lean=lean_partial:for:one_sided

More for · 2
2005 · cited by 0
... show up widely , as well , in shared nonlinguistic cognitive systems ... word use at least , a great amount of our usage is of words for referents ... Universal ( cross - cultural and cross - language ) patterns have been used ... The Encyclopedia of Social Measurement captures the data, techniques, theories, designs, applications, histories, and implications of assigning numerical values to social phenomena. Responding to growing demands for transdisciplinary descriptions of quantitative and qualitative techniques, measurement, sampling, and statistical methods, it will increase the proficiency of everyone who gathers and analyzes data. Covering all core social science disciplines, the 300+ articles of the Encyclopedia of Social Measurement not only present a comprehensive summary of observational frameworks and mathematical models, but also offer tools, background information, qualitative methods, and guidelines for structuring the research process. Articles include examples and applications of research strategies Covering all core social science disciplines, the 300+ articles of the Encyclopedia of Social Measurement not only present a comprehensive summary of observational frameworks and mathematical models, but also offer tools, background information, qualitative methods, and guidelines for structuring the research process. Articles include examples and applications of research strategies and techniques, highlighting multidisciplinary options for observing social phenomena. The alphabetical arrangement of the articles, their glossaries and cross-references, and the volumes' detailed index will encourage exploration across the social sciences. Descriptions of important data sets and case studies will help readers understand resources they can often instantly access. Also available online via ScienceDirect - featuring extensive browsing, searching, and internal cross-referencing between articles in the work, plus dynamic linking to journal articles and abstract databases, making navigation flexible and easy. For more information, pricing options and availability visit www.info.sciencedirect.com. Readers are provided with references for further information Eleven substantive sections delineate social sciences and the research processes they follow to measure and provide new knowledge on a wide range of topics Authors are prominent scholars and methodologists from all social science fields Within each of the sections important components of quantitative and qualitative research methods are dissected and illustrated with examples from diverse fields of study Actual research experiences provide useful examples Dari dalam buku Ada 1 halaman yang cocok dengan Word frequencies exhibit universal cross-lingual patterns dalam buku ini Halaman 365 Ke mana bagian lainnya buku ini? Her work has brought innovation to measurement of diverse topics, including criminal career patterns, gender bias, racial disparity, insider trading, judicial decision-making, and theories of crime and justice. She is probably best known for her research aimed at understanding and improving juvenile justice system policies and procedures. Much of her applied research to assist state and local criminal justice agencies with policy evaluation and reform has been supported by state and federal funding. Her work has appeared in leading criminology journals and several edited books. Tracy, she is preparing a follow-up volume to Continuity & Discontinuity in Criminal Careers (1997, Plenum) about life course patterns of the 14,000 females in the 1958 Philadelphia Birth Cohort Study, including details about self-reported victimization, offending, and other personal experiences obtained from a survey administered to a sample of the cohort in their early 20s. She also is working with the Dallas County Juvenile Department.
cited by 0
improvements in parsing through the use of word and phrase frequencies provide more compelling evidence): ... might be normalized occurrence frequencies of particular words (or word classes) and punctuation. Especially ... relations, and features with broader scope such as word frequencies or document class. Apart from this difference ...
Everything we examined (3)
This check searched the claim as stated. It did not run a separate search for evidence against it.
  1. Encyclopedia of Social Measurementreferenceno side taken
  2. Cross-lingual Embeddings Reveal Universal and Lineage-Specific Patterns in Grammatical Gender Assignmentpeer-reviewedno side taken
  3. Computational Linguisticsreferenceno side taken
The paper trail · every fact has a biography
held for human review08 Aug 2026
This receipt carries no identity, shared or not. Sharing publishes your connection to it, not your data.
Check your own claim
Challenge the receipt
trust me, bro: win the argument, pass the class, survive peer review.
This receipt is an automated verdict against our published method · not an opinion about any author or publication.
Terms · Privacy · How verdicts work · Dispute this receipt