Task: N/A
Release Date: 9/15/2026
Format: WAV, MPG, EAF, MD, CSV
Size: 1.54 GB
This collection contains nine Iskonawa recordings from 2013 documenting collective oral traditions and conversations. The materials include narratives about the origin of the Iskonawa people, cremation practices (two versions), the spirit tree, the hualo, the stingy ones, and the peanut paucar, as well as open conversations among several speakers. Unlike the personal narratives, these recordings predominantly involve multi-speaker interaction, with three speakers (S01, S02, and S03) participating together in most recordings, making the collection particularly valuable for the study of interaction and turn-taking. Recordings were made by Roberto Zariquiey Biondi in the Native Communities of Chachibai and Callería and in San José de Yarinacocha, Ucayali, Peru. Transcription, translation, and ELAN processing were also carried out by Roberto Zariquiey Biondi. The research team included Roberto Zariquiey, José Antonio Mazzotti, and Carolina Rodríguez.
Licensing
Licencia Chana 2.0 — Licence for Peruvian Indigenous Language Documentation Collections
https://github.com/rzariquiey/licencia-chanaRestrictions/Special Constraints
This dataset is released under the Licencia Chana 2.0, not under a Creative Commons licence. Download is direct and unrestricted, but use is subject to the following terms. PERMITTED WITHOUT FURTHER AUTHORISATION - Academic research and publication, with attribution. - Teaching and educational use, with attribution. - Language revitalisation work by the source communities and by organisations working with them. - Evaluating and benchmarking existing speech or language models, with attribution. - Non-production academic machine-learning experimentation, provided that (a) the resulting model is not deployed in production, (b) model weights are not released, and (c) any results published cite this collection. REQUIRES PRIOR WRITTEN AUTHORISATION FROM THE DONOR - Training models intended for production or public release. - Release or distribution of model weights derived from this material. - Any commercial use. - Redistribution of the dataset, in whole or in part, on any other platform. ACCESS IS GRANTED CASE BY CASE, WITH INDIVIDUAL REVIEW. Requesters must identify themselves, state their institutional affiliation, and describe the intended use in specific terms. The donor reviews each request individually and may decline without stating a reason. This collection has been placed at the most restrictive level because of the nature of its content, not because consent is lacking. All speakers gave their authorisation. The restriction exists because this is material that should circulate with context and accompaniment rather than by open download. Recommended Traditional Knowledge labels (Local Contexts): TK Community Use Only, TK Non-Commercial, TK Attribution. ATTRIBUTION Cite as: Zariquiey, Roberto (donor). Iskonawa: Narraciones-Conversaciones. Archivo Digital de Lenguas Peruanas, Pontificia Universidad Católica del Perú, Lima. https://hdl.handle.net/20.500.12534/4CTKGE RIGHT OF WITHDRAWAL Speakers retain the right to withdraw their recordings from this collection at any time. This is the principal reason a Creative Commons licence was not used: CC licences are irrevocable and cannot accommodate this commitment. This collection includes accounts of funerary practice, including two versions of the cremation narrative, and oral tradition of ritual value. Contact for authorisations: [email protected]
Forbidden Usage
- You agree not to attempt to determine the identity of the pseudonymised speakers in this dataset, nor to link the speaker codes (S01, S02...) to real individuals. - Any attempt to clone, synthesise or imitate the voices of the speakers in this dataset is forbidden. - Training models intended for production deployment or public release is forbidden without prior written authorisation from the donor. This includes releasing model weights trained wholly or partly on this material. - Commercial use of any kind is forbidden without prior written authorisation. - Redistribution of this dataset, in whole or in part, on any other platform or in any other repository is forbidden without prior written authorisation. - Use of this material in ways that misrepresent, decontextualise or commercially exploit the cultural knowledge it contains is forbidden. Note on machine learning: this is not a blanket prohibition on ML research. Evaluating and benchmarking existing models on this data is permitted, and so is non-production academic experimentation, provided the model is not deployed, the weights are not released, and the results cite the collection. What requires authorisation is production training and weight release. The reason is that the consent forms signed before approximately 2020 do not mention AI training and cannot reasonably be read as covering it; the communities have not yet been consulted on this point. Contact for authorisations: [email protected]
Ethical Review
All recordings in this collection were made with the informed consent of the speakers, obtained in the field at the time of recording. Signed consent forms are held by the donor at the Pontificia Universidad Católica del Perú and are available to the MDC team on request. Speakers are identified by code (S01, S02...) rather than by name, in file names, in the annotation tiers of the ELAN files, and in all accompanying documentation. Contributors are acknowledged collectively in each collection's documentation, without linking any name to any specific recording or code. A small number of individuals appear under their own names because they asked to be credited as authors of the material; this is stated explicitly in the documentation of the collections concerned. Material judged sensitive was withheld from deposit entirely rather than published under restriction. Collections involving funerary practice, medicinal knowledge, ritual knowledge, or extended life histories of identifiable individuals are published under restricted access with individual review of each request. One point remains open and is declared here rather than glossed over: the consent forms signed before approximately 2020 do not mention the use of recordings for training artificial intelligence models. They cannot reasonably be read as covering it. The communities have not yet been consulted on this specific question. Until that consultation takes place, production model training is not authorised, and the licence reflects this.
Intended Use
The central texts of Iskonawa oral tradition, including the origin narrative, together with multi-party conversation. Intended for research on interaction and turn-taking in a critically endangered language, and for preservation with context.