Skip to Main Content
Article navigation
Purpose

This paper aims to provide a context for Brazilian Portuguese language documentation and its data collection to establish linguistic repositories from a sociolinguistic overview.

Design/methodology/approach

The main sociolinguistic projects that have generated collections of Brazilian Portuguese language data are presented.

Findings

The comparison with another situation of repositories (seed vaults) and with the accounting concept of assets is evocated to map the challenges to be overcome in proposing a standardized and professional language repository to host the collections of linguistic data arising from the reported projects and others, in the accordance with the principles of the open science movement.

Originality/value

Thinking about the sustainability of projects to build linguistic documentation repositories, partnerships with the information technology area, or even with private companies, could minimize problems of obsolescence and safeguarding of data, by promoting the circulation and automation of analysis through natural language processing algorithms. These planning actions may help to promote the longevity of the linguistic documentation repositories of Brazilian sociolinguistic research.

Licensed re-use rights only
You do not currently have access to this content.
Don't already have an account? Register

Purchased this content as a guest? Enter your email address to restore access.

Please enter valid email address.
Email address must be 94 characters or fewer.
Pay-Per-View Access
$41.00
Rental

or Create an Account

Close Modal
Close Modal