3. Producción

Browse

Search Results

Now showing 1 - 10 of 14
  • Some of the metrics are blocked by your 
    Item type:Publication,
    Using lexical language models to detect borrowings in monolingual wordlists
    (Public Library of Science, 2020-12-01)
    Lexical borrowing, the transfer of words from one language to another, is one of the most frequent processes in language evolution. In order to detect borrowings, linguists make use of various strategies, combining evidence from various sources. Despite the increasing popularity of computational approaches in comparative linguistics, automated approaches to lexical borrowing detection are still in their infancy, disregarding many aspects of the evidence that is routinely considered by human experts. One example for this kind of evidence are phonological and phonotactic clues that are especially useful for the detection of recent borrowings that have not yet been adapted to the structure of their recipient languages. In this study, we test how these clues can be exploited in automated frameworks for borrowing detection. By modeling phonology and phonotactics with the support of Support Vector Machines, Markov models, and recurrent neural networks, we propose a framework for the supervised detection of borrowings in mono-lingual wordlists. Based on a substantially revised dataset in which lexical borrowings have been thoroughly annotated for 41 different languages from different families, featuring a large typological diversity, we use these models to conduct a series of experiments to investigate their performance in mono-lingual borrowing detection. While the general results appear largely unsatisfying at a first glance, further tests show that the performance of our models improves with increasing amounts of attested borrowings and in those cases where most borrowings were introduced by one donor language alone. Our results show that phonological and phonotactic clues derived from monolingual language data alone are often not sufficient to detect borrowings when using them in isolation. Based on our detailed findings, however, we express hope that they could prove to be useful in integrated approaches that take multi-lingual information into account.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    How to commentate a soccer match in Shipibo-Konibo (Pano)?
    (University of Campinas, 2020-02-12)
    The present paper lists and illustrates eleven strategies that are systematically used by Shipibo-Konibo speakers in order to comment live soccer matches in the context of an indigenous soccer cup informally called “Mundialito Shipibo”. We argue that these lexical, morphosyntactic and discursive strategies can be classified into three types according to their function: iconic strategies, which attempt to present the information more vividly (onomatopoeic forms and ideophones, reduplications, parallel structures and hearsay-quotatives); emotional strategies, which are used by soccer commentators to express their emotions and their feelings (interjections, player-directed speech and diminutives); and proximity strategies, which bring the speech closer to the Shipibo-Konibo audience (lexical Shipibo-Konibo innovations, vocatives, evidential access configurations and code alternations). These various strategies are crucial for understanding the new social dynamics that the Shipibo-Konibo language is getting into as a consequence of becoming an urban language, and are clearly creating a new speech genre. The new social uses that the Shipibo-Konibo people are giving to their language and the features that the language is developing in this new social context are crucial to understand the future of Shipibo-Konibo and other minority languages in Peru.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    On the von Neumann entropy of language networks: applications to cross-linguistic comparisons
    (IOP Publishing Ltd, 2021-12-01)
    Words are not isolated entities within a language. In this paper, we measure the number of choices transmitted in natural language by means of the von Neumann entropy of language networks. This quantity, introduced in Quantum Information accounts, provides a detailed characterization of network complexities. The simulations are based on a large parallel corpus of 362 languages across 55 linguistic families (focusing on the sub-sample of 85 languages from the Americas). With this, we constructed language networks as a simple way to describe word connectivity patterns for each language. We studied several aspects of the von Neumann entropy of language networks. First, we discovered large groups of languages with low average degree and high von Neumann entropy. The results suggested also that large von Neumann entropy is associated with word entropy (as a proxy for morphological complexity), and is inversely related to degree regularity. This means that there are pressures at play that keep a balance between word morphological complexity and patterns of connections between words. We suggested also a strong influence of functional words on low von Neumann entropy languages. Our approach is thus a simple network-based contribution to establish cross-linguistic language comparisons from textual data.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    Grambank reveals the importance of genealogical constraints on linguistic diversity and highlights the impact of language loss
    (Center for Open Science, 2022-12-04)
    While global patterns of human genetic diversity are increasingly well characterized, the diversity of human languages remains less systematically described. Here we outline the Grambank database. With over 400,000 data points and 2,400 languages, Grambank is the largest comparative grammatical database available. The comprehensiveness of Grambank allows us to quantify the relative effects of genealogical inheritance and geographic proximity on the structural diversity of the world's languages, evaluate constraints on linguistic diversity, and identify the world's most unusual languages. An analysis of the consequences of language loss reveals that the reduction in diversity will be strikingly uneven across the major linguistic regions of the world. Without sustained efforts to document and revitalize endangered languages, our linguistic window into human history, cognition and culture will be seriously fragmented.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    Untangling the evolution of body-part terminology in Pano: conservative versus innovative traits in body-part lexicalization
    (Royal Society, 2022-12-09)
    Although language-family specific traits which do not find direct counterparts outside a given language family are usually ignored in quantitative phylogenetic studies, scholars have made ample use of them in qualitative investigations, revealing their potential for identifying language relationships. An example of such a family specific trait are body-part expressions in Pano languages, which are often lexicalized forms, composed of bound roots (also called body-part prefixes in the literature) and non-productive derivative morphemes (called here body-part formatives). We use various statistical methods to demonstrate that whereas body-part roots are generally conservative, body-part formatives exhibit diverse chronologies and are often the result of recent and parallel innovations. In line with this, the phylogenetic structure of body-part roots projects the major branches of the family, while formatives are highly non-tree-like. Beyond its contribution to the phylogenetic analysis of Pano languages, this study provides significative insights into the role of grammatical innovations for language classification, the origin of morphological complexity in the Amazon and the phylogenetic signal of specific grammatical traits in language families.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    Realization of the Tripartite Alignment in Amahuaca: Synchronic Analysis and Diachronic Explanation
    (Universidad Nacional Mayor de San Marcos, 2023-09-27)
    The native language Amahuaca (Panoan family, Peru) features a cross-linguistically unusual tripartite case-marking system whereby the subjects of intransitive and transitive verbs are signaled by means of the enclitics =x and =n, respectively, whereas the objects do not bear any morphological marking (Sparing-Chávez, 2007/2012; Clem, 2019a; among others). Another salient property of this system is the diverse realizations that nominals display when they host overt case-markers; this is, precisely, the primary focus of the present article. While these manifestations might initially appear erratic, our analysis reveals that the realizations of the marked nominals are highly predictable, especially when we posit the presence of a latent consonant that emerges phonetically in the context of =x and =n. From a diachronic perspective, we argue that this allomorphic variabil-ity can be linked to a rule proposed by Shell (1975), which is believed to have affected trisyllabic items in the protolanguage.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    Language Classification in Western Amazonia
    (University of Campinas, 2023-02-16)
    The languages of the Pano and Takana families exhibit a considerable number of lexical and structural affinities that cannot be ascribed to mere chance and are not readily detectable instances of borrowing. After the comparative studies by Key (1968) and Girard (1971) the proposal of a genetic relationship between these two families was generally accepted (e.g. Loos 1973, 2005; Suárez 1973; Kaufman 1990; Campbell 1997). Without solid argumentation, however, this classification was later put into question (Fabre 1998; Loos 1999; Fleck 2013) and, even today, there is no full consensus as to whether the observed similarities are due to genetic inheritance or long-term language contact. The present paper offers lexical and grammatical evidence in support of the hypothesis that Pano and Takana are genetically connected. Comparing for the first time what can be considered Proto-Pano and Proto-Takana reconstructions, it is shown that 18 of the 40 items in the basic vocabulary list proposed by the Automated Similarity Judgment Program (asjp) (Holman et al. 2008) might be cognate; this includes 9 body-part terms. Also, a set of alleged grammatical cognates are assembled, and shared constructions involving motion verbal morphology, intransive and transitive auxiliaries, transitivity harmony restrictions, and switch-reference are discussed.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    Grambank Reveals the Importance of Genealogical Constraints on Linguistic Diversity and Highlights the Impact of Language Loss
    (American Association for the Advancement of Science, 2023-04-21)
    While global patterns of human genetic diversity are increasingly well characterized, the diversity of human languages remains less systematically described. Here, we outline the Grambank database. With over 400,000 data points and 2400 languages, Grambank is the largest comparative grammatical database available. The comprehensiveness of Grambank allows us to quantify the relative effects of genealogical inheritance and geographic proximity on the structural diversity of the world’s languages, evaluate constraints on linguistic diversity, and identify the world’s most unusual languages. An analysis of the consequences of language loss reveals that the reduction in diversity will be strikingly uneven across the major linguistic regions of the world. Without sustained efforts to document and revitalize endangered languages, our linguistic window into human history, cognition, and culture will be seriously fragmented.
  • Some of the metrics are blocked by your 
    Item type:Publication,
    The poetics of the journey in the traditional songs of the Kakataibo people
    (University of Alicante, 2025-01-01)
    El presente artículo se propone contribuir al conocimiento sobre el arte verbal del pueblo kakataibo, el cual se manifiesta en una compleja red de formas cantadas de lenguaje que se usaron tradicionalmente por miembros de este pueblo en distintos espacios sociales e íntimos. Este sistema establecía una distinción bastante nítida entre géneros que podían ser cantados por hombres y géneros que podían ser cantados por mujeres. Nuestro estudio se centra en un motivo que se repite en varios de las canciones que forman parte de nuestro corpus musical del pueblo kakataibo: el viaje. Mostramos aquí las estrategias poéticas y discursivas empleadas por hombres y mujeres para hablar del viaje: el uso de colores y de metáforas, por ejemplo, pero también la construcción de un yo poético radicalmente diferente para cada caso. El hombre kakataibo canta de sus viajes, se muestra fuerte, valiente y siempre guerrero. Es él quien viaja. La mujer kakataibo, por otra parte, se muestra frágil. Llora recordando a quienes partieron. Ella no viaja, pero se quiebra al ver partir a sus familiares. Se trata de una dualidad que se complementa, un mismo viaje visto desde dos perspectivas distintas, cantadas desde dos yo poéticos distintos. El hombre indestructible y la mujer frágil son, en realidad, dos personajes idealizados que se van construyendo a través de los versos.
      2
  • Some of the metrics are blocked by your 
    Item type:Publication,
    The forgotten mid vowels
    (John Benjamins Publishing Company, 2026-07-10)
    This study re-examines the phonological inventory of Proto-Pano, challenging the prevailing view that it had only four oral vowels. Through a comparative analysis of 23 contemporary Pano languages/varieties, we reconstruct a Proto-Pano system with two additional mid vowels, proposing a six-vowel inventory. This finding diverges from all previous reconstructions, which overlooked historically significant distributional differences between mid vowels in Northern Pano languages, Kakataibo, and Kaxarari. We argue that Pano mid vowels exhibit both conservative and innovative traits, questioning the assumption that all Pano mid vowels are innovations. Our analysis highlights the role of merger in shaping Pano phonological diversity, supporting a six-vowel reconstruction as a more accurate reflection of the comparative data. This revised perspective underscores the importance of carefully analyzing sound correspondence patterns and offers new insights into the historical evolution and internal classification of the Pano language family. Additionally, it calls for a thorough revision of previous Proto-Pano reconstructions and opens promising avenues for future research on the Pano-Takana stock.