Sign in

Thomas Pellard

@thomaspellard.bsky.social
425 followers 343 following 194 posts

CNRS Research scientist 🇫🇷 historical-comparative linguistics, archaeolinguistics, phylolinguistics, geolinguistics, endangered languages, languages of Japan

PostsRepliesMedia
Reposted by Thomas Pellard
Sampsa Holopainen @sampsaholopainen.bsky.social · 25/09/2026
This paper (by me, Terhi Honkola and Samuli Helama) on possible connections between the 4.2 ka climatic event and the early spread of the Uralic languages was published last week.
sciencedirect.com
Extreme events and the spread of Uralic languages
In this paper we present novel aspects concerning the possible connection between the 4.2 ka climatic event and the early spread of Uralic languages f…
11913
Reposted by Thomas Pellard
JLLQ @physiologos.bsky.social · 24/09/2026
Pas facile de répondre en quelques secondes à des questions impromptues, qui mériteraient quelques heures de développement… www.arte.tv/fr/videos/13...
arte.tv
Il a recensé les mythes de plus 1300 peuples à travers le monde ! - 28 minutes (24/09/2026) - Regarder l’émission complète | ARTE
L'anthropologue Jean-Loïc Le Quellec publie le "Dictionnaire critique de mythologie" aux éditions du CNRS. Dans cet ouvrage de 1553 pages, il explore une diversité de mythes de tous les continents, is...
043
Reposted by Thomas Pellard
Dahlem Center for Linguistics @dclberlin.bsky.social · 21/09/2026
#OTD 268 years ago, Antoine Isaac Silvestre de Sacy (1758–1838) was born 🎉 He was an expert in Arabic, Syriac, and Persian, and the first French linguist to attempt to decipher the Egyptian hieroglyphic texts on the Rosetta Stone. #LinguisticBirthdays #Histlx
087
Reposted by Thomas Pellard
Johann-Mattis List @lingulist.de · 21/09/2026
First version of the newly created Database of Cross-Linguistic Partial Colexifications (CLIPS) has now been published online. Common work with Katja Bocklage, Thanasis Georgakopoulos, Robert Forkel, and myself clips.digling.org
Logo of the CLIPS database with a directed network combining three nodes.
157
Reposted by Thomas Pellard
Jeffrey Shallit 🇺🇦 @shallit.bsky.social · 18/09/2026
A new risk for mathematicians: a colleague reports that immediately after they posted an abstract of their upcoming talk online, a student elsewhere used AI to derive proofs of the stated results and then posted it on the arxiv. My colleague hadn't posted to arxiv yet. Scummy. Watch out.
16219
Thomas Pellard @thomaspellard.bsky.social · 19/09/2026
「現代の言語の科学としての日本語学にとって、語源論は、その中心的な研究対象や範疇に入らない」🤔 語源論は現代の日琉歴史比較言語学の中心的な研究対象に入っている
130
Thomas Pellard @thomaspellard.bsky.social · 17/09/2026
New preprint, on the etymology of Japanese mina ‘all’! 📰 Reflecting on the origin of Japanese ‘all’ ➡️ hal.science/hal-05718204
hal.science
Making sure you're not a bot!
175
Reposted by Thomas Pellard
Diachronica journal @diachronica.bsky.social · 16/09/2026
Reconstructing pharyngealized vowels for Qieyun Division II in Middle Chinese benjamins.com/catalog/dia....
benjamins.com
Reconstructing pharyngealized vowels for Qieyun Division II in Middle Chinese
It is widely accepted that Middle Chinese Division II, a syllable category in Middle Chinese, originates from the Old Chinese medial *-r-, yet its reconstruction remains debated. Existing proposals as...
031
Reposted by Thomas Pellard
Chris Buckley @chrisbuckley.bsky.social · 15/09/2026
On Thursday 17 Sept I will be giving a talk at the Max Planck Institute for the history of science (MPIWG) in Berlin, title is "Signal to Noise: the Transmission of Traditional Weaving Cultures". It's part of the IASSRT symposium, and open to anyone to attend
Mother and daughter warping a loom, in Flores, Indonesia
294
Reposted by Thomas Pellard
Martin Haspelmath @haspelmath.bsky.social · 11/09/2026
Version 1.1. of the Grammaticon is now online. I keep working on it, but finding good definitions is often hard. For 138 of the 516 concepts, the definition is still preliminary (indicated by [[...]]). Suggestions are welcome! grammaticon.clld.org/concepts
0139
Reposted by Thomas Pellard
Dahlem Center for Linguistics @dclberlin.bsky.social · 14/09/2026
#OTD 235 years ago, Franz Bopp (1791–1867) was born 🎂 He was a specialist in Sanskrit and a pioneer of the comparative analysis of Indo-European grammatical systems, aiming to reconstruct their common ancestor. At least three streets in Germany are named after him. #LinguisticBirthdays #Histlx
0134
Reposted by Thomas Pellard
Martin Haspelmath @haspelmath.bsky.social · 14/09/2026
Can typological grammar-surveying be automatized? Four years ago, I was at a workshop in Uppsala, where a "grammar-reading machine" was discussed (spraakbanken.gu.se/en/projects/...). Now it seems that LLMs can do the job (arxiv.org/abs/2609.07791), so what is left for human typologists?
152
Reposted by Thomas Pellard
Robin Ryder @robinryder.bsky.social · 12/09/2026
A strong call to remember what Mathematics research is actually about. Thanks you for writing this. (I signed.)
In recent months, the success of AI in solving major mathematical problems has made headlines even outside mathematical circles. But solving problems is only a tool and proxy for achieving the primary goal of conceptual understanding and insight. Forgetting this in the world of AI may turn the tool against the primary goal. Indeed, the mass production at faster and faster pace of "true/false" statements could destroy fertile ground instead of breathing life into new ideas.
064
Reposted by Thomas Pellard
Carlos Scheidegger @cscheid.net · 10/09/2026
quarto-hub is really coming together, y'all! I'm so excited to share some of this with you at conf next week. Check out our upcoming collaborative wysiwyg editor working on a fully-styled Quarto 2 page:
56514
Reposted by Thomas Pellard
rgyalrong.bsky.social @rgyalrong.bsky.social · 05/09/2026
New article by David Goldstein: www.cambridge.org/core/journal...
cambridge.org
Articles are inversely associated with inflectional case in Indo-European | Language | Cambridge Core
Articles are inversely associated with inflectional case in Indo-European
052
Reposted by Thomas Pellard
Center for the Human Past (CHP) @humanpastcenter.bsky.social · 01/09/2026
The application for the interdisciplinary course “Reducing the Gap Between Linguistics and Genetics” is open until 7 September. Please spread the word and notify any MA/PhD students you think might benefit! Link: sites.utu.fi/humandiversi... @haam-community.bsky.social @cschlebu.bsky.social
083
Reposted by Thomas Pellard
Centre de recherches linguistiques sur l'Asie orientale @crlao-ling.bsky.social · 28/08/2026
🔜13th International Conference of the European Asociation of Chinese Linguistics (EACL-13) 📆2-4 September 2026 📍Centre de colloques, Campus Condorcet, Aubervilliers crlao.cnrs.fr/event/13th-i...
crlao.cnrs.fr
The 13th International Conference of the European Association of Chinese Linguistics
Dear colleagues, it is with great pleasure that we invite you to participate in the 13th International Conference of the European Association of Chinese Linguistics (EACL-13), held at the Centre de Colloques, Campus Condorcet at Aubervilliers from Sept. 2nd to Sept. 4th, 2026, to which many…
032
Reposted by Thomas Pellard
rgyalrong.bsky.social @rgyalrong.bsky.social · 27/08/2026
panchr.hypotheses.org/4103
panchr.hypotheses.org
Existe-t-il un étymon désignant l’armoise en proto-sino-tibétain ?
Le Shijing atteste un nombre considérable de type d'armoises (genre Artemisia), et l'une d'entre elle, 蘩 bjon (Artemisia sieversiana selon Pan Fujun 2003), me semblait à première vue être un candidat ...
022
Reposted by Thomas Pellard
Marijn van Putten @phdnix.bsky.social · 27/08/2026
In 2025 @bnuyaminim.bsky.social and me taught a course on Comparative Semitics at the Leiden Summer School. After some cleaning up, we've decided to upload the syllabus and release it under a CC license for you and yours to peruse through. doi.org/10.17613/b4z...
27434
Reposted by Thomas Pellard
Dahlem Center for Linguistics @dclberlin.bsky.social · 26/08/2026
#OTD 131 years ago, Jerzy Kuryłowicz (1895–1978) was born 🎉 In his academic career, he focused on Indo-European linguistics. He is best known for his Laws of Analogy, but he also contributed to the study of case, grammaticalization, and to the laryngeal theory. #LinguisticBirthdays #Histlx
0101
Reposted by Thomas Pellard
laserhedvig @laserhedvig.bsky.social · 26/08/2026
A New Language Game has Entered The Villa: Glottle! glottle.app Try to guess the target language and along the way you get clues from its position in a language family tree and its geographical position. #languages #linguistics
21811
Reposted by Thomas Pellard
Mark Hudson @barbarianniche.bsky.social · 26/08/2026
New paper www.frontiersin.org/journals/hum...
frontiersin.org
Frontiers | Hunter-gatherer experiences of colonialism do not explain subsequent patterns of socio-economic development
How modern colonialism impacted socio-economic growth before and after independence is the subject of an extensive literature. Employing ideas from neoclassi...
021
Reposted by Thomas Pellard
Casey Dunn @caseywdunn.bsky.social · 25/08/2026
My new book, Phylogenetic Biology, is now published. Available online for free at dunnlab.org/phylogenetic... . Physical copies can be purchased at shop.lightningsource.com/b/085?params... or your favorite bookstore.
12315160
Reposted by Thomas Pellard
laserhedvig @laserhedvig.bsky.social · 24/08/2026
Next week in Berlin (31st August, 2nd September), MPI-directors Russell Gray and Steven Levinson will be giving public talks. More details in flyers below.
163
Reposted by Thomas Pellard
Chris Buckley @chrisbuckley.bsky.social · 24/08/2026
Attending European Archaeology conf in Athens? I will be giving two talks. The first, with Guillaume Jacques, on Thurs 27th at 8.45am (session 191, room 640) on a novel method for analysing spindle whorl data, and its application to trace the spread of weaving in East Asia ...
a map showing the emergence and spread of weaving technologies in East Asia
174
Reposted by Thomas Pellard
🐕‍🦺 @welt.bsky.social · 19/08/2026
github.com/anti-tail/De...
github.com
GitHub - anti-tail/DeltaDrops
Contribute to anti-tail/DeltaDrops development by creating an account on GitHub.
121
Reposted by Thomas Pellard
Johann-Mattis List @lingulist.hcommons.social.ap.brid.gy · 19/08/2026
In our August contribution to our blog Computer-Assisted Language Comparison in Practice #CALCiP, we discuss The Ambiguity of the Colexification Term (with Katja Bocklage and Thanasis Georgakopoulos) calc.hypotheses.org/9547 doi.org/10.15475/calcip.2026.2.2
calc.hypotheses.org
The Ambiguity of the Colexification Term
The term colexification has become an important concept in lexical typology in particular and linguistic typology in general. Surprisingly, it is not consistenly used by all scholars in the field. Instead, two main interpretations have emerged, a broad version that uses colexification as a cover term for polysemy, vagueness, and homophony, and a narrow version that restricts the term to polysemy and vagueness. We discuss the ambiguity and reflect about the usefulness of introducing alternative terms instead. ## 1 Introduction The introduction of the term _colexification_ to the field of lexical typology by François (2008), can be seen as marking a turning point in the field of lexical typology. Before, lexical typology was a niche discipline pursued by a small number of scholars, concentrating on very specific domains of the lexicon of human languages, such as color, kinship, and body. After the introduction of the term in a widely recognized volume on lexical typology by Vanhove (2008), devoted to the role that polysemy plays in investigating semantic shift, both the available data and the methodology by which the data could be analyzed increased drastically. In 2010, Cysouw (2010) showed, how recurrent polysemies can be visualized and analyzed as a network. In 2011, Urban (2011) presented the idea that certain directional tendencies of semantic change can be inferred from the existence of examples in which semantic similarities are overtly marked in individual language varieties. In 2014, the _Database of Cross-Linguistic Colexifications_ (CLICS, https://clics.clld.org) was published (List et al., 2014), building on Cysouws’ idea of polysemy networks which were derived from a collection of 300 comparative wordlists that were automatically searched for _colexifications_ , that is, cases where different concepts in the same language were expressed by identical word forms. With these new resources and ideas to visualize and explore semantic relations in large-scale accounts of cross-linguistic data, lexical typology grew into a popular subfield of linguistic typology (Koptjevskaja-Tamm et al. 2007, Koptjevskaja-Tamm et al. 2015, Sjöberg et al. forthcoming). While previous studies had compared how the lexicons of human languages organize meaning in different ways, had relied on case studies and small-scale examples involving hand-selected languages, it was now possible to study meaning organization through the lens of a large part of the world’s languages with the help of technically and theoretically considerably simple means. ## **2 From Polysemy to Colexification** François originally defined _colexification_ as a relation that holds for the senses that are expressed by identical word forms in one and the same language. > A given language is said to colexify two functionally distinct senses if, and only if, it can associate them with the same lexical form. (François, 2008, p. 170) From this definition, it is clear that colexification is intended as a _cover term_ _for the previously established terminology around_ __polysemy__ _,___vagueness__ _, and_ __homophony__. In linguistic terminology, there are several terms that can be used to _specific_ cases where two senses are associated with the same lexical form, such as _polysemy_ and _vagueness_ , referring to senses associated with the _same_ form, and _homophony_ , referring to senses associated with different forms that came to sound alike at some point in the history of a language. From the perspective of linguistic terminology, the term was useful in two regards. Since, on the one hand, it replaces largely _explanative terminology_ with _descriptive terminology_ _(see Jacques & List, 2019, p. 141 on this distinction)_, it facilitates, on the other hand, the investigation of lexical semantics across languages, since it shifts the burden of proof from the stage of data collection to the stage of data analysis. Explanative terminology (one could also say “diagnostic terminology”) does not only describe a phenomenon, but also tries to explain it at the same time. Polysemy is usually _explained_ as the result of semantic change when individual forms acquire more senses, while homophony is _explained_ as the result of sound change leading to phonological merger in individual forms of a language. Determining that two senses are colexified in a given language does not explain _why_ this is the case, but merely notes this as a fact inferred from empirical observations. This has drastic consequences for the investigation of lexical semantics across languages. While any cross-linguistic study on polysemy, vagueness, or homophony, would have had to distinguish these phenomena beforehand, reducing the study to the overall phenomenon of distinct senses being expressed by the same form enabled large-scale investigations based on standardized comparative wordlist collections. In such a facilitated workflow, one would no longer have to search individual languages for instances of polysemy in a fixed set of meanings, but could rather search for _colexifications_ , that is, cases where two or more concepts in a comparative wordlist are expressed by identical translational equivalents. How these cases could be _classified_ would be a second step, where early research immediately showed that cases of polysemy and vagueness tend to recur across different language families, while cases of homophony tend to be restricted to individual languages and individual language families (List et al., 2013). As a result, the distinction between polysemy and vagueness on the one hand and homophony on the other hand, can be readily established _after_ colexification patterns have been assembled for a sufficiently large number of languages and language families. This procedure was the basic idea behind the CLICS database, which provides an automated procedure to harvest colexification patterns from cross-linguistic wordlist collections that was modified and improved across the database’s four different major versions (List et al., 2014, 2018; Rzymski et al., 2020; Tjuka et al., 2025). When using colexification as a cover term for polysemy, vagueness, and homophony, CLICS assembles individual examples from different languages of all these phenomena and later allows to single out cases of homophony by focusing on the most frequently observed _patterns_. ## **3 Intended Definition of Colexification** Recent discussions at the workshop on _Semantic Shifts and the Dynamics of Lexification Patterns_ organized by Alexandre François, Anna Zalizniak, and Anna Smirnitskaya as part of the _16th International Conference of the Association for Linguistic Typology in Lyon_ (July 1-3, 2026) made it clear that the term _colexification_ is used in two different ways by linguists from different backgrounds (or “schools”). On the one hand, there are linguists who work either in the tradition of François’ original work or in the context of the _Database of Semantic Shifts_ (https://datsemshift.ru, Zalizniak et al., 2023), a large-scale catalogue of various attested kinds of semantic shifts that have been collected for more than 20 years now directly from the linguistic literature (Zalizniak, 2018). On the other hand, there are linguists working in the context of the CLICS database. The former tradition restricts the term _colexification_ to cases of polysemy and vagueness, excluding cases of homophony from the definition. In the context of the work of the latter tradition, _colexification_ was always a cover term for _all_ cases where one form has multiple senses in a given language. Given that the tradition restricting colexification to polysemy and vagueness is represented by the scholar who introduced the term colexification, there are good reasons to follow the narrower definition of colexification, even if the original text introducing the term does not explicitly exclude homophony as one of the aspects covered by the term. Even in the case of the CLICS database, which was profiting so much from the broader version of colexification in order to justify its workflow, the narrow version of colexification might be useful and even help to clarify past studies. From these studies on emotion concepts (Jackson et al., 2019), body part terms (Tjuka, 2024; Tjuka et al., 2024), perception and cognition (Georgakopoulos et al. 2022), and lexical creativity (Brochhagen et al., 2023), it is quite clear that the main intention of the database is to _infer_ colexifications in the narrow sense, filtering out cases of homophony by rather simple quantitative means. However, treating colexification as a mere cover term for vagueness and polysemy means at the same time, that the two major advantages of the term, the shift from explanative to descriptive terminology and the facilitation that this shift provides for quantitative approaches, would vanish, at least in parts. It would also mean that terminological frameworks that attempt to view colexification in the larger context of _coexpression_ would have to be revised, especially since cases of _cogrammification_ as defined by Haspelmath (2023) also do not ask _how_ a pattern emerged, but rather conclude that a pattern can be _determined_ from the data. Additionally, it may be questionable whether it is actually possible to strictly distinguish between homophony on the one hand and polysemy and vagueness on the other hand in all cases (consider the ambiguity of English _ear_ , as referring to a body part or the part of a plant, which is often seen as some kind of polysemy, although it is the result of phonological merger). Thus, despite François’ original intention, one may argue that the broad version of colexification is actually in line with the established practice in linguistic typology. As often, when it comes to terminology, it seems that the field will have to live with some compromise solutions that emerge from the research practice. Abandoning the broad definition of colexification would require us to employ a new term, which may be seen as a minor problem, but given the conflict with the practice in general discussions on coexpression and semantic maps, we would be forced to revise the terminology in this area more generally and broadly. Retaining the broad definition for colexification as the norm has the disadvantage that in most scholars’ _practice_ , homophony is essentially ignored, no matter whether they see colexification in its broad or narrow version. It seems for the time being, that it is best to leave everything as is, while being prepared to mention the ongoing conflict with the terminology explicitly, where needed. ## **4 Beyond Colexification** While François agreed in the discussion about the term _colexification_ at the workshop at the ALT conference, that the original paper (François, 2008) does _not_ explicitly exclude homophony from the definition, reading the original paper with the revised narrow definition in mind clarifies those parts that go beyond _strict colexification_ proper. > In particular, “strict colexification” (same lexeme in synchrony) should be carefully distinguished from “loose colexification” (covering all other cases mentioned here). (François, 2008, p. 171) _Strict colexification_ is contrasted with _loose colexification_. Under the latter term François (2008) subsumes a variety of other, more or less similar, relations between concepts mediated through their word forms. Among these are _diachronic semantic change_ , in which a word form _w_ expresses a concept A at a point in time _t1_ but refers to concept B at a later point in time _t2_ (at which stage concept A may either still be attested or have been lost). Cases involving diachronic semantic change are comparable to strict colexification in that the concepts relate through an identical word form; unlike strict colexification, however, these word forms belong to different historical stages. Owing to this similarity Georgakopoulos & Polis (2021) do not treat the two cases as entirely distinct, but instead introduced the term _diachronic strict colexification_ , which captures both the diachronic dimension of the phenomenon and the fact that the same form is involved. While the broad definition of colexification may be seen as problematic here, since cases in which word forms overlap in parts, are the norm rather than the exceptions in natural languages, even if the word forms involved do not display any etymological connections, it is clear from the examples given by François that colexification is used in the _narrow_ sense. This means that François does not consider cases where two word forms show some coincidental overlap in some parts, but rather presupposes that loose colexifications can only be determined if one can prove the historical unity of the shared material in question. While it seems basically convincing to restrict colexification along these lines, recent research devoted to so-called _partial colexifications_ , that is, colexification patterns involving _parts_ of the linguistic form (List et al., 2022), seem to show again, that the notion of colexification as a descriptive term has concrete advantages for computational studies. Thus, List (2023) showed how the idea of _synchronic asymmetries_ in _overt marking_ proposed by Urban (2011) can be operationalized in a quantitative framework and applied to larger collections of comparative wordlists by searching automatically for those cases in which one word occurs at the beginning or the end of another word in a given language. Since these substring relations are called _affixes_ in computer science, List proposes to term these relations _affix colexifications_ , as a specific case of _partial colexification_ , emphasizing that the inference of these relations across different forms in a given language does not necessarily reflect valid semantic relations, but showing that potentially valid forms can again be inferred by filtering _patterns_ across multiple languages and language families. Partial colexification can be seen as more restrictive than loose colexification, given that the latter refers to etymological relations, while the former explicitly points to the synchronic identity of _parts_ of the linguistic form in a given language. That cross-linguistically recurring affix colexifications can give hints to valid and potentially interesting semantic relations is further illustrated by a study of Bocklage et al. (2025). In an empirical evaluation, the authors show that filtering affix colexifications by frequency of occurrence across language families can indeed help minimizing the number of false positives. Abandoning the term _partial colexification_ due to its conflict in coverage with the term _loose colexification_ may again seem to be the easiest way to deal with the terminological problems here. But again, we must observe that first, the definition of loose colexifications does not explicitly exclude cases of coincidence in form overlap. Second, abandoning the broad colexification definition would yield conflicts with the general coexpression terminology in linguistic typology, given that coexpression cannot rule out coincidence as easily as lexical typology can rule out homophony. As a result, it may be useful to emphasize that the terms _loose colexification_ and _partial colexification_ are not compatible, given that they reflect different scholarly traditions, the former being at least in part explanative (definitely diagnostic), and the latter being deliberately descriptive. ## **5 Conclusion** From what has been discussed so far, it will probably be clear that the terminology with respect to and around the term _colexification_ has made some infelicitous turns so far. As a result, we are – yet another time – faced with terminological traditions in our field that display certain conflicts among scholars, differences in opinions, and also sheer misunderstandings. On the one hand, the young history that the term _colexification_ experienced so far, can be seen as representative for typical terminology struggles in science. On the other hand, one should not ignore the momentum of serendipity reflected in the fate of the term _colexification_ , given that the reinterpretation of the partially explanative term as a purely descriptive term has given scholars the chance to develop simple but very effective quantitative approaches to lexical typology. Future research will show if additional solutions can be found. For now, all we can recommend is that scholars who work with colexifications make clear in which tradition they allocate their terminology. **References** Bocklage, K., Georgakopoulos, T., Dam, K. P. van, Ciucci, L., Blum, F., Kučerová, A., Rubehn, A., Stephen, A., Snee, D., & List, J.-M. (2025). Testing the Potential of Automatically Inferred Affix Colexifications for Linguistic Typology. in _Humanities Commons_ (pp. 1–35). https://doi.org/10.17613/a06m1-c9939 Brochhagen, T., Boleda, G., Gualdoni, E., & Xu, Y. (2023). From language development to language evolution: A unified view of human lexical creativity. _Science_ , _381_(6656), 431–436. https://doi.org/10.1126/science.ade7981 Cysouw, M. (2010). Drawing Networks from Recurrent Polysemies. Comment on “polysemous qualities and universal networks” by loïc-michel perrin (2010). _Linguistic Discovery_ , _8_(1), 281–285. François, A. (2008). Semantic maps and the typology of colexification: intertwining polysemous networks across languages. in M. Vanhove (Ed.), _From polysemy to semantic change_ (pp. 163–215). Benjamins. Georgakopoulos, T. and Polis, S. (2021): Lexical diachronic semantic maps. Mapping the evolution of time-related lexemes. _Journal of Historical Linguistics_ 11(3). 367-420. https://doi.org/10.1075/jhl.19018.geo Georgakopoulos, T., Grossman, E., Nikolaev, D., and S Polis (2022): Universal and macro-areal patterns in the lexicon: A case-study in the perception-cognition domain. _Linguistic Typology_ 26(2): 439–487. https://doi.org/10.1515/lingty-2021-2088 Haspelmath, M. (2023). Coexpression and synexpression patterns across languages: comparative concepts and possible explanations. _Frontiers in Psychology_ , _14_ , 1–12. https://doi.org/10.3389/fpsyg.2023.1236853 Jackson, J. C., Watts, J., Henry, T. R., List, J.-M., Mucha, P. J., Forkel, R., Greenhill, S. J., Gray, R. D., & Lindquist, K. (2019). Emotion semantics show both cultural variation and universal structure. _Science_ , _366_(6472), 1517–1522. https://doi.org/10.1126/science.aaw8160 __ Jacques, G., & List, J.-M. (2019). Save the trees: Why we need tree models in linguistic reconstruction (and when we should apply them). _Journal of Historical Linguistics_ , _9_(1), 128–166. https://doi.org/10.1075/jhl.17008.mat Koptjevskaja-Tamm, M., Vanhove, M., and Koch, P. (2007). Typological approaches to lexical semantics. Linguistic Typology, 11(1), 159–185. https://doi.org/10.1515/LINGTY.2007.013 Koptjevskaja-Tamm, M., Rakhilina, E., and M. Vanhove (2015): The Semantics of Lexical Typology. In Nick Riemer (ed.), _The Routledge Handbook of Semantics_ (pp. 434–454). London: Routledge. List, J.-M. (2023). Inference of partial colexifications from multilingual wordlists. _Frontiers in Psychology_ , _14_(1156540), 1–10. https://doi.org/10.3389/fpsyg.2023.1156540 List, J.-M., Forkel, R., Greenhill, S. J., Rzymski, C., Englisch, J., & Gray, R. D. (2022). Lexibank, A public repository of standardized wordlists with computed phonological and lexical features. _Scientific Data_ , _9_(316), 1–31. https://doi.org/10.1038/s41597-022-01432-0 __ List, J.-M., Greenhill, S. J., Anderson, C., Mayer, T., Tresoldi, T., & Forkel, R. (2018). CLICS². An improved database of cross-linguistic colexifications assembling lexical data with help of cross-linguistic data formats. _Linguistic Typology_ , _22_(2), 277–306. https://doi.org/10.1515/lingty-2018-0010 List, J.-M., Mayer, T., Terhalle, A., & Urban, M. (2014). _CLICS: Database of Cross-Linguistic Colexifications. Version 1.0_ (version 1.0.0). Forschungszentrum Deutscher Sprachatlas. http://clics.lingpy.org List, J.-M., Terhalle, A., & Urban, M. (2013). Using network approaches to enhance the analysis of cross-linguistic polysemies. _Proceedings of the 10th International Conference on Computational Semantics – Short Papers_ , 347–353. Rzymski, C., Tresoldi, T., Greenhill, S., Wu, M.-S., Schweikhard, N. E., Koptjevskaja-Tamm, M., Gast, V., Bodt, T. A., Hantgan, A., Kaiping, G. A., Chang, S., Lai, Y., Morozova, N., Arjava, H., Hübler, N., Koile, E., Pepper, S., Proos, M., Epps, B. V., … List, J.-M. (2020). The Database of Cross-Linguistic Colexifications, reproducible analysis of cross- linguistic polysemies. _Scientific Data_ , _7_(13), 1–12. https://doi.org/10.1038/s41597-019-0341-x Sjöberg, A., Georgakopoulos, T., and Koptjevskaja-Tamm, M. (forthcoming): Typology and universals of word meaning. In D. Geeraerts & D. Glynn (Eds.), _The Cambridge handbook of lexical semantics_. Cambridge University Press. Tjuka, A. (2024). Objects as human bodies: cross-linguistic colexifications between words for body parts and objects. _Linguistic Typology_ , _0_(0). https://doi.org/10.1515/lingty-2023-0032 Tjuka, A., Forkel, R., & List, J.-M. (2024). Universal and cultural factors shape body part vocabularies. _Scientific Reports_ , _14_(10486), 1–12. https://doi.org/10.1038/s41598-024-61140-0 Tjuka, A., Forkel, R., Rzymski, C., & List, J.-M. (2025). Advancing the Database of Cross-Linguistic Colexifications with New Workflows and Data. _Proceedings of the 16th International Conference on Computational Semantics_ , 1–15. https://aclanthology.org/2025.iwcs-main.1 Urban, M. (2011). Asymmetries in overt marking and directionality in semantic change. _Journal of Historical Linguistics_ , _1_(1), 3–47. Vanhove, M. (Ed.). (2008). _From polysemy to semantic change_. Benjamins. Zalizniak, A. A. (2018). The Catalogue of Semantic Shifts: 20 years later. _Russian Journal of Linguistics_ , _22_(4), 770–787. https://doi.org/10.22363/2312-9182-2018-22-4-770-787 Zalizniak, A. A., Smirnitskaya, A., Russo, M., Mikhailova, T., Bobrik, M., Gruntov, I., Orlova, M., Bibaeva, M., & Voronov, M. (2023). _Database of Semantic Shifts [Version from 20/12/2023)_. Institute of Linguistics at the Russian Academy of Sciences. https://datsemshift.ru/ **Cite this article as:** Johann-Mattis List, Katja Bocklage, and Thanasis Georgakopoulos (2026): “The ambiguity of the colexification term” in _Computer-Assisted Language Comparison in Practice_ , 9.2: 85-92 [first published on 19/08/2026], URL: https://calc.hypotheses.org/9547, DOI: 10.15475/calcip.2026.2.2. **Download the article as PDF:** calcip-09-2-2.pdf **Copyright information** : This is an open access article distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited. **Funding Information** : This project has received funding from the European Research Council (ERC) under the European Union’s Horizon Europe research and innovation programme (Grant agreement No. 101044282). The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript. Johann-Mattis List Seit Anfang 2023 leite ich den Lehrstuhl für Multilinguale Computerlinguistik in Passau. In meiner Forschung nehme ich generell einen datenbasierten,… Read more View all posts Katja Bocklage View all posts Thanasis Georgakopoulos View all posts * * * The text only may be used under licence Creative Commons Attribution 4.0 International. All other elements (illustrations, imported files) are “All rights reserved”, unless otherwise stated. * * * OpenEdition suggests that you cite this post as follows: Johann-Mattis List, Katja Bocklage, Thanasis Georgakopoulos (August 19, 2026). The Ambiguity of the Colexification Term. _Computer-Assisted Language Comparison in Practice_. Retrieved August 19, 2026 from https://calc.hypotheses.org/9547 * * * * * * * *
022
Reposted by Thomas Pellard
Yale University @yale.edu · 12/08/2026
Thousands of languages have disappeared throughout human history. New research is shedding light on why. A study coauthored by Yale linguist Claire Bowern finds that widespread language loss began not with European colonialism, but roughly a thousand years earlier.
bit.ly
Study uncovers lost ‘golden age' of languages
A new study coauthored by Yale linguist Claire Bowern suggests that tens of thousands of languages were spoken between 1,000 and 3,000 years ago.
0157
Thomas Pellard @thomaspellard.bsky.social · 18/08/2026
New preprint! (Yes, another one) Exploring the impact of inference method, character coding scheme, and cognacy criterion on phylolinguistic inference: The case of the Japonic language family osf.io/preprints/so...
osf.io
OSF
053
Thomas Pellard @thomaspellard.bsky.social · 18/08/2026
New preprint!Japanese has effectively only two accent patterns: A quantitative and information-theoretic approach to lexical predictability osf.io/preprints/so...
osf.io
OSF
074
Reposted by Thomas Pellard
Daniel Lakens @lakens.bsky.social · 15/08/2026
New Blog post: Which Data Repository Should you Use? In light of OSF closing down, I compare Zenodo, Dataverse, ResearchBox, PsychArchive, and local repositories on six important dimensions. If you want to know which to pick: It depends! daniellakens.blogspot.com/2026/08/whic...
daniellakens.blogspot.com
Which Data Repository Should You Use?
The Center for Open Science has announced that from November 16, 2026, no new projects can be created on the Open Science Framework. After F...
4197111
Reposted by Thomas Pellard
Nabil Hathout @nabilhathout.bsky.social · 13/08/2026
New work on how to transpose the Paradigm Cell Filling Problem to derivation. Work done with Basilio Calderone, Franck Sajous and Fiammetta Namer, to appear in Word Structure (EUP). hal.science/hal-05714785
Abstract. A computational exploration of the Paradigm Cell Filling Problem in derivation
141
Reposted by Thomas Pellard
Aki Vehtari @avehtari.bsky.social · 13/08/2026
All four books I've co-authored are freely available online for non-commercial use: - Bayesian Workflow at avehtari.github.io/Bayesian-Wor... Links to other three books are in the quoted post 👇 (too many books to fit in one post!)
avehtari.github.io
Bayesian Workflow book: Website – Bayesian Workflow book
Website for the Bayesian Workflow book by Gelman, Vehtari, McElreath, et al. — case studies, code, and exercises in R and Stan.
8283133
Reposted by Thomas Pellard
Václav Hrnčíř @vaclavhrncir.bsky.social · 13/08/2026
I'm pleased that our chapter on Genes, Identity and Material Culture has also been included in this book. link.springer.com/chapter/10.1...
link.springer.com
Genes, Identity, and Material Culture: In Search of a Unifying Theory
This chapter discusses the cultural identity of past societies, an issue that has traditionally been addressed through the concept of archaeological cultures - discrete evolutionary units defined by t...
1102
Reposted by Thomas Pellard
vincentholst.bsky.social @vincentholst.bsky.social · 13/08/2026
In 2023, @nature.com published 'Papers and patents are becoming less disruptive over time', receiving world-wide media attention. Our Matters Arising, published after a 32 month delay (more on that soon), shows that the reported decline can largely be attributed to dataset artefacts. 🧵
The average CD_5 index per year for Web of Science. The original Park et al. decline (top curve) becomes essentially flat (bottom curve) when removing papers with CD_5=1. Those papers largely correspond to dataset artefacts.
9351173
Reposted by Thomas Pellard
César Castellvi @cesar-castellvi.bsky.social · 11/08/2026
À l’approche de la sortie d’un livre consacré aux kawaraban, j’ai publié dans The Conversation un bref billet sur ces ancêtres populaires des journaux modernes dans le Japon d’Edo. Et si le sujet vous intéresse, la suite arrive très bientôt 😊 theconversation.com/informer-sur...
theconversation.com
Informer sur les catastrophes à l’époque du Japon féodal : le rôle des « kawaraban »
Au Japon, le duo information et divertissement n’a pas attendu l’arrivée des journaux modernes. Les « kawaraban » ont longtemps permis de couvrir toutes sortes d’actualités et de divertir malgré la ce...
054
Reposted by Thomas Pellard
rgyalrong.bsky.social @rgyalrong.bsky.social · 11/08/2026
I just published a new article, "Harvesting knives in Neolithic East Asia", coauthored with Chris Stevens, Kan Yu-Chun and Dorian Fuller: doi.org/10.1016/j.qu...
doi.org
Redirecting
123
Reposted by Thomas Pellard
Brian Nosek @briannosek.bsky.social · 11/08/2026
Over the next several months, we will phase out some popular OSF features. These changes do not affect current public content, which remains safe and accessible. We have prepared support material to help users adapt. I am very sorry for the disruption to your work. Please read this post for details.
cos.io
OSF Changes | Center for Open Science
We are preparing substantial changes to OSF that will reduce its functionality, focus OSF on its unique strengths, and move toward an integrated model with complementary services. As part of that shif...
10120113
Reposted by Thomas Pellard
Dahlem Center for Linguistics @dclberlin.bsky.social · 07/08/2026
#OTD 180 years ago, Hermann Paul (1846–1921) was born 🎂 A medievalist focusing on Middle High German, lexicologist, lexicographer, and member of the Neogrammarian school, he is probably best known as the author of "Principien der Sprachgeschichte", published in 1880. #LinguisticBirthdays #Histlx
0144
Reposted by Thomas Pellard
Mark Hudson @barbarianniche.bsky.social · 04/08/2026
First drafted this for the Oxford Handbook of the Archaeology of Japan & Korea almost 10 years ago. It needs updating! But I wanted to post the draft to receive any feedback about the 'medieval Ainu diaspora', a concept inspired by Viking studies. papers.ssrn.com/sol3/papers....
papers.ssrn.com
<p><span>The Okhotsk Culture and the Formation of the Medieval Ainu Diaspora</span></p>
Across Eurasia the Bronze Age saw the beginning of new patterns of migration and interaction. Long-distance trade in metals, new transport, food and textile tec
063
Reposted by Thomas Pellard
Simon J. Greenhill @simongreenhill.bsky.social · 04/08/2026
What do we know about the language extinction rate? doi.org/10.1017/ext.2026.10019
doi.org
What do we know about the language extinction rate? | Cambridge Prisms: Extinction | Cambridge Core
What do we know about the language extinction rate?
1139
Reposted by Thomas Pellard
Markˣ @shjsat.bsky.social · 30/07/2026
I'm cross-posting @haspelmath.bsky.social "Gradient in grammatical structure of indigenous languages reflects pathway of human expansion in the Americas" Link to paper: www.nature.com/articles/s41...
Facebook post by Martin Haspelmath discussing Urban & Guzman-Naranjo's paper
124
Reposted by Thomas Pellard
Alexei Drummond @alexeidrummond.bsky.social · 28/07/2026
Get excited :) PhyloSpec is coming! A shared standard + modelling language for phylogenetic models. Write a model once, run it across engines like BEAST X, BEAST 2.8 & RevBayes. Easy and reproducible. Built by @tochsner.bsky.social & collaborators (ETH Zürich, Auckland, LMU). phylospec.com
phylospec.com
PhyloSpec
A standardized way to describe phylogenetic model components, common assumptions, and best practices in the field of phylogenetics.
14926
Reposted by Thomas Pellard
波鴻漫錄 || Sven Osterkamp @schrift-sprache.bsky.social · 23/07/2026
For the proceedings of the previous /gʁafematik/ conference (23–25 October 2024), see now (in OA 🙏): Part I: www.fluxus-editions.fr/gla11.php Part II: www.fluxus-editions.fr/gla12.php
fluxus-editions.fr
Fluxus Editions
Grapholinguistics and Its Applications
023
Reposted by Thomas Pellard
Damián Blasi @damianblasi.bsky.social · 23/07/2026
How many languages have existed over the Holocene—and what does that reveal about the design space of languages and cultures? Now out in @science.org www.science.org/eprint/QDZNY....
science.org
The rise and fall of language diversity through the Holocene
Characterizing the factors that have shaped linguistic diversity is fundamental for understanding human history, culture, and cognition. In this study, we combined statistical and social computational...
213254
Reposted by Thomas Pellard
Centre de recherches linguistiques sur l'Asie orientale @crlao-ling.bsky.social · 21/07/2026
La Lettre d'information du #CRLAO n°31, Juillet/Août 2026 est en ligne ! Pour connaître l'actualité scientifique et d'autres informations sur la vie du laboratoire : mailchi.mp/1b310944a053...
mailchi.mp
La Lettre du CRLAO n° 31, Juillet/Août 2026
011
Reposted by Thomas Pellard
Johann-Mattis List @lingulist.de · 20/07/2026
Just appeared in the Journal of Language Evolution, by David Snee, Luca Ciucci, and myself: "Variation in language phylogenies may result from variation in concept translation" doi.org/10.1093/jole...
doi.org
Variation in language phylogenies may result from variation in concept translation
Abstract. Phylogenetic reconstruction in historical linguistics now typically relies on cognates sets assembled from multilingual wordlists. While more and
071
Reposted by Thomas Pellard
Christophe Nahon @chnahon.bsky.social · 18/07/2026
Le Centre national de la recherche scientifique (CNRS) annonce l’accélération de sa sortie de certains logiciels extra-européens pour gagner en souveraineté numérique et réduire ses coûts IT. www.zdnet.fr/actualites/l...
zdnet.fr
Le CNRS sort de Zoom et Microsoft Exchange, et bientôt Sharepoint - ZDNET
Le Centre national de la recherche scientifique annonce l’accélération de sa sortie de certains logiciels extra-européens pour gagner en souveraineté numérique et réduire ses coûts IT.
033
Reposted by Thomas Pellard
Caley Orr @caleyorr.bsky.social · 18/07/2026
🏺🧪
nature.com
Population expansion of early Upper Palaeolithic hunter-gatherers triggered cultural complexity on Paleo-Honshu Island, Japan - Nature Communications
Early Upper Palaeolithic stone tools appear in Japan around 38 thousand years ago and are unique from contemporaneous tools found elsewhere in eastern Eurasia. Here, the authors present an absolute ch...
1219
Reposted by Thomas Pellard
Chris Buckley @chrisbuckley.bsky.social · 18/07/2026
Edge-ground stone tools and other interesting features, not associated with agriculture, appeared in Japan around 36kya
two ground stone axe blades
031