User Information
Structure
The basic ontology of RWAAI is the 13 recognised branches of the Austroasiatic family. Some of the branches are further divided into lower subfamily branches, but the level of granularity is dependent on areal expertise. The collections are grouped under the relevant language names, with each individual collection typically named after the creator. The collections will also be browsable by secondary ontologies created from the catalogue metadata, including ISO code, branch name, depositor's name, and geographic location.
There are 4 access levels for archived materials:
RWAAI is committed to creating a fully accessible and usable resource. There is however no binding obligation for depositors to provide general access to their collections. The depositor will set the access rights to resources in accordance with ethical and legal concerns specific to the collection. This may result in materials being permanently unavailable for public access if their distribution could potentially cause harm or distress to the participants. In recognition of the diversity of materials contained in a corpus, each individual item can be designated an individual access specification. The depositor may also choose to apply other conditions on the access or use of materials at their own discretion. It is the depositor’s responsibility to ensure that they have sought the appropriate permissions from the creators of the materials in deciding on access levels. RWAAI bears no responsibility for the dissemination of materials by individual depositors.
Open: No login or registration is required
Available to registered users: Send a request to the curator to register.
Access needs to be requested: By using the "Request Access" function, you can apply for access to these materials as a registered user. The depositor will be asked whether or not access can be granted to you.
Closed: These materials are currently not accessible for reasons determined by the owner.
Linked: Materials that are linked from other archives have an icon with no color. Consult the source archive for access levels.
The Corpus Server
For the end user (e.g. a researcher who wants to upload and manage data sets or a student who wants permission to access data) the corpus server mainly has two ways of interaction. The first one is the IMDI browser, which is used for browsing and searching through the various data sets. The metadata (a.k.a. information about data) is open for anyone to browse and search.
- Browse RWAAI's collections: https://archive.humlab.lu.se/
The second one is LAMUS, which is mainly used by researchers and administrators to upload and manage data sets.
- LAMUS: https://corpora.humlab.lu.se/jkc/lamus/
There is a link to the LAMUS manual on the LAMUS main page.
Registration, deposition, access
Questions regarding registration, deposition of data and access to data in the RWAAI branch should be directed to the RWAAI curator.
Working with the Corpus Server and technical issues
If you experience technical issues or if you have questions regarding your account or workflow (e.g. “How do I create IMDI-files?”), please contact the Corpus Manager.
Current Users
Your current accounts should maintain their access levels from the previous server version. If this is not the case, send an e-mail to the RWAAI curator. Please make sure to include your user name and details regarding what is not working, what part of the corpus tree you should have access to etc.
Access levels
Your current accounts should maintain their access levels from the previous server version. If this is not the case, send an e-mail to the Corpus Manager. Please make sure to include your user name and details regarding what is not working, what part of the corpus tree you should have access to etc.New Users
You can browse the corpus tree and the metadata without registering but if you do wish to register please send a request to the RWAAI curator.
Note: Being registered will not automatically permit access to data. Some of the data sets might be open access for registered users but in general access restrictions to data is decided by each respective owner/researcher. Data is by default ONLY available to the person uploading the data and/or to users who already have download access to the node/s in question. Access requests for specific resources should be directed to the RWAAI curator.
Metadata and IMDI-files
Metadata is information describing your data (data about your data). It can include but is not limited to location where the data was collected, language/s, participants etc. This information is required in order to add data and will be linked to the corresponding data file/s. Metadata is what will help you keep track of your data and to find it when needed.
Please remember that each researcher will have to decide whether a piece of information should be included in the metadata set or not due to sensitivity/legal issues (e.g. if you include the actual names of consultants in your metadata, these will also be visible to anyone browsing the corpus tree).
Arbil is necessary to create the metadata files a.k.a. IMDI-files (a standardized metadata-format, stored in xml) that you will upload together with your data. In short, no data can be added to the archive without corresponding metadata.
The latest version and manuals can be downloaded for free from The Language Archive at the MPI (Windows/Mac/Linux):Please note that while Arbil is a desktop application, it is written in Java and consequently requires Java to be installed on your system.
