Search | re3data.org

Filter

Reset all

Subjects

Content Types

Countries

AID systems

API

Certificates

Data access

Data access restrictions

Database access

Database access restrictions

Database licenses

Data licenses

Data upload

Data upload restrictions

Enhanced publication

Institution responsibility type

Institution type

Keywords

Metadata standards

PID systems

Provider types

Quality management

Repository languages

Software

Syndications

Repository types

Versioning

Toogle short help

* at the end of a keyword allows wildcard searches
" quotes can be used for searching phrases
+ represents an AND search (default)
| represents an OR search
- represents a NOT operation
( and ) implies priority
~N after a word specifies the desired edit distance (fuzziness)
~N after a phrase specifies the desired slop amount

← Previous
1 (current)
2
Next →

Found 28 result(s)

The World Atlas of Language Structures

WALS

Subject(s)

Content type(s)

Country

Germany

The World Atlas of Language Structures (WALS) is a large database of structural (phonological, grammatical, lexical) properties of languages gathered from descriptive materials (such as reference grammars) by a team of 55 authors (many of them the leading authorities on the subject).

Kikapu

University of the Western Cape Institutional Research Data Respository

Subject(s)

Content type(s)

Country

South Africa

The University of the Western Cape (UWC) uses Figshare for Institutions for their institutional research data repository. It is called Kikapu, and serves as a repository for storing and disseminating research data.

Språkbanken Text

Språkbanken

Subject(s)

Content type(s)

Country

Språkbanken was established in 1975 as a national center located in the Faculty of Arts, University of Gothenburg. Allén's groundbreaking corpus linguistic research resulted in the creation of one of the first large electronic text corpora in another language than English, with one million words of newspaper text. The task of Språkbanken is to collect, develop, and store (Swedish) text corpora, and to make linguistic data extracted from the corpora available to researchers and to the public.

Tools for Second Language Speech Research and Teaching

SLA Speech Tools

Subject(s)

Content type(s)

Country

This website constitutes a repository of tools and resources for researchers and teachers that are interested in second language speech acquisition and pronunciation teaching in diverse educational contexts. If you are a RESEARCHER in the field of second language acquisition (SLA), here you will find a wide range of validated tools that may be useful for your individual differences, SLA or L2 speech studies. If you are a passionate second language pronunciation TEACHER interested in communicative methods, here you will be able to download several carefully designed explicit instruction, communicative form-focused activities and pronunciation-based tasks that are ready to be used in your classroom

IANUS Datenportal

Digitale Forschungsdaten aus Archäologie & Altertumswissenschaften

Subject(s)

Content type(s)

Country

Germany

IANUS is a DFG-funded project to set up a national research data center for archeology and ancient science in Germany.

ARCHE

A Resource Centre for the HumanitiEs

Subject(s)

Content type(s)

Country

ARCHE (A Resource Centre for the HumanitiEs) is a service aimed at offering stable and persistent hosting as well as dissemination of digital research data and resources for the Austrian humanities community. ARCHE welcomes data from all humanities fields. ARCHE is the successor of the Language Resources Portal (LRP) and acts as Austria’s connection point to the European network of CLARIN Centres for language resources.

bonndata

Subject(s)

Content type(s)

Country

Germany

bonndata is the institutional, FAIR-aligned and curated, cross-disciplinary research data repository for the publication of research data for all researchers at the University of Bonn. The repository is fully embedded into the University IT and Data Center and curated by the Research Data Service Center (https://www.forschungsdaten.uni-bonn.de/en). The software that bonndata is based on is the open source software Dataverse (https://dataverse.org)

CLARIN-ERIC

Common Language Resources and Technology Infrastructure - European Research Infrastructure Consortium

Subject(s)

Content type(s)

Country

CLARIN is a European Research Infrastructure for the Humanities and Social Sciences, focusing on language resources (data and tools). It is being implemented and constantly improved at leading institutions in a large and growing number of European countries, aiming at improving Europe's multi-linguality competence. CLARIN provides several services, such as access to language data and tools to analyze data, and offers to deposit research data, as well as direct access to knowledge about relevant topics in relation to (research on and with) language resources. The main tool is the 'Virtual Language Observatory' providing metadata and access to the different national CLARIN centers and their data.

Pacific and Regional Archive for Digital Sources in Endangered Cultures

PARADISEC

Subject(s)

Content type(s)

Country

Australia

PARADISEC (the Pacific And Regional Archive for Digital Sources in Endangered Cultures) offers a facility for digital conservation and access to endangered materials from all over the world. Our research group has developed models to ensure that the archive can provide access to interested communities, and conforms with emerging international standards for digital archiving. We have established a framework for accessioning, cataloguing and digitising audio, text and visual material, and preserving digital copies. The primary focus of this initial stage is safe preservation of material that would otherwise be lost, especially field tapes from the 1950s and 1960s.

datastore by Universität Münster

Subject(s)

Content type(s)

Country

Germany

datastore is the cross-domain research data repository of the University Münster (Germany). In datastore, scientific members of the University Münster can publish their research data following the FAIR principles, including the assignment of a DOI for each dataset as a persistent identifier.

O2 Repositori UOC - Dades

Subject(s)

Content type(s)

Country

Spain

O2 is the UOC's institutional repository. Section 'Dades' contains primary data accompanying documents published in the Research and Institutional communities.

CLARIN-LT

CLARIN-LT Repository

Subject(s)

Content type(s)

Country

Lithuania became a full member of CLARIN ERIC in January of 2015 and soon CLARIN-LT consortium was founded by three partner universities: Vytautas Magnus University, Kaunas Technology University and Vilnius University. The main goal of the consortium is to become a CLARIN B centre, which will be able to serve language users in Lithuania and Europe for storing and accessing language resources.

melbourne.figshare.com

University of Melbourne data repository

Subject(s)

Content type(s)

Country

melbourne.figshare.com is a specialised service that has been tailored according to specific needs and requirements of the University and our community of researchers. The service offered at the University is free to use, provides 100GB of data, and stores all data on the University's storage system.

D-PLACE

Database of Places, Language, Culture and Environment

Subject(s)

Content type(s)

Country

D-PLACE contains cultural, linguistic, environmental and geographic information for over 1400 human ‘societies’. A ‘society’ in D-PLACE represents a group of people in a particular locality, who often share a language and cultural identity. All cultural descriptions are tagged with the date to which they refer and with the ethnographic sources that provided the descriptions. The majority of the cultural descriptions in D-PLACE are based on ethnographic work carried out in the 19th and early-20th centuries (pre-1950).

University of Manchester figshare

Subject(s)

Content type(s)

Country

United Kingdom

The University selected figshare as a general purpose research data repository to enable researchers to share research data, facilitate open research practices and meet the evolving requirements of research funders and academic publishers. This is a public-facing platform for researchers to share their data and build, over time, a comprehensive representation of the research done at the University across all faculties and disciplines.

Open Research Data Online

ORDO

Subject(s)

Content type(s)

Country

United Kingdom

The figshare service for The Open University was launched in 2016 and allows researchers to store, share and publish research data. It helps the research data to be accessible by storing metadata alongside datasets. Additionally, every uploaded item receives a Digital Object Identifier (DOI), which allows the data to be citable and sustainable. If there are any ethical or copyright concerns about publishing a certain dataset, it is possible to publish the metadata associated with the dataset to help discoverability while sharing the data itself via a private channel through manual approval.

CLARIN-IS

CLARIN á Íslandi

Subject(s)

Content type(s)

Country

Iceland joined CLARIN ERIC on February 1st, 2020, after having been an observer since November 2018. The Ministry of Education, Science and Culture assigned The Árni Magnússon Institute for Icelandic Studies the role of leading partner in the Icelandic National Consortium and appointed Professor Emeritus Eiríkur Rögnvaldsson as National Coordinator, later replaced by Starkaður Barkarson, a project manager at The Árni Magnússon Institute. Most of the relevant institutions participate in the CLARIN-IS National Consortium. The Árni Magnússon Institute has already established a Metadata Providing Centre (CLARIN C-Centre) which hosts metadata for Icelandic language resources and makes them available through the Virtual Language Observatory. The aim is to establish a Service Providing Centre (CLARIN B-Centre) which will provide both service and access to resources and knowledge.

Repositório de Dados de Pesquisa Unifesp

Unifesp Data Repository

Subject(s)

Content type(s)

Country

Brazil

The Unifesp Research Data Repository is a platform for storing, preserving and accessing research data for the institution's academic community.

Sprachatlas Baden-Württemberg

Sprachatlas BW

Subject(s)

Content type(s)

Country

Germany

The speaking language atlas gives a multimedia impression of the dialects of the state Baden-Württemberg in Germany. The maps of the Speaking Language Atlas of Baden-Württemberg are based on two databases: Südwestdeutschen Sprachatlas (SSA) and the Sprachatlas von Nord Baden-Württemberg (SNBW). The dialect recordings that form the basis for the maps were carried out at the SSA between 1974 and 1986, but at the SNBW between 2009 and 2012. For the southern part, this means that the maps may present a state of affairs that is no longer valid today.

Multimodal Learning Corpus Exchange

MULCE

Subject(s)

Content type(s)

Country

Mulce (MUltimodal contextualized Learner Corpus Exchange) is a research project supported by the National Research Agency (ANR programme: "Corpus and Tools in the Humanities", ANR-06-CORP-006). A teaching corpus (LETEC - Learning and Teaching Corpora) combines a systematic and structured data set, particularly of interactional data, and traces left by a training course experimentation, conducted partially or completely online and completed by additional technical, human, pedagogical and scientific information to enable the data to be analysed in context.

ILC-CNR for CLARIN-IT repository

ILC4CLARIN

Subject(s)

Content type(s)

Country

ILC-CNR for CLARIN-IT repository is a library for linguistic data and tools. Including: Text Processing and Computational Philology; Natural Language Processing and Knowledge Extraction; Resources, Standards and Infrastructures; Computational Models of Language Usage. The studies carried out within each area are highly interdisciplinary and involve different professional skills and expertises that extend across the disciplines of Linguistics, Computational Linguistics, Computer Science and Bio-Engineering.

TROLLing

Tromsø Repository of Language and Linguistics

Subject(s)

Content type(s)

Country

The Tromsø Repository of Language and Linguistics (TROLLing) is a FAIR-aligned repository of linguistic data and statistical code. The archive is open access, which means that all information is available to everyone. All data are accompanied by searchable metadata that identify the researchers, the languages and linguistic phenomena involved, the statistical methods applied, and scholarly publications based on the data (where relevant). Linguists worldwide are invited to deposit data and statistical code used in their linguistic research. TROLLing is a special collection within DataverseNO (http://doi.org/10.17616/R3TV17), and C Centre within CLARIN (Common Language Resources and Technology Infrastructure, a networked federation of European data repositories; http://www.clarin.eu/), and harvested by their Virtual Language Observatory (VLO; https://vlo.clarin.eu/).

Informatics Research Data Repository

IDR

Subject(s)

Content type(s)

Country

Japan

The Informatics Research Data Repository is a Japanese data repository that collects data on disciplines within informatics. Such sub-categories are things like consumerism and information diffusion. The primary data within these data sets is from experiments run by IDR on how one group is linked to another.

Das Deutsche Referenzkorpus

DeReKo

Subject(s)

Content type(s)

Country

The project is set up in order to improve the infrastructure for text-based linguistic research and development by building a huge, automatically annotated German text corpus and the corresponding tools for corpus annotation and exploitation. DeReKo constitutes the largest linguistically motivated collection of contemporary German texts, contains fictional, scientific and newspaper texts, as well as several other text types, contains only licenced texts, is encoded with rich meta-textual information, is fully annotated morphosyntactically (three concurrent annotations), is continually expanded, with a focus on size and stratification of data, may be analyzed free of charge via the query system COSMAS II, serves as a 'primordial sample' from which users may draw specialized sub-samples (socalled 'virtual corpora') to represent the language domain they wish to investigate. !!! Access to data of Das Deutsche Referenzkorpus is also provided by: IDS Repository https://www.re3data.org/repository/r3d100010382 !!!

CLARIN.SI repository

Slovenian CLARIN repository

Subject(s)

Content type(s)

Country

CLARIN.SI is the Slovenian node of the European CLARIN (Common Language Resources and Technology Infrastructure) Centers. The CLARIN.SI repository is hosted at the Jožef Stefan Institute and offers long-term preservation of deposited linguistic resources, along with their descriptive metadata. The integration of the repository with the CLARIN infrastructure gives the deposited resources wide exposure, so that they can be known, used and further developed beyond the lifetime of the projects in which they were produced. Among the resources currently available in the CLARIN.SI repository are the multilingual MULTEXT-East resources, the CC version of Slovenian reference corpus Gigafida, the morphological lexicon Sloleks, the IMP corpora and lexicons of historical Slovenian, as well as many other resources for a variety of languages. Furthermore, several REST-based web services are provided for different corpus-linguistic and NLP tasks.

← Previous
1 (current)
2
Next →

Current projects
EOSC FAIR-IMPACT

re3data COREF

To the extent possible under law, re3data.org has waived all copyright and related or neighboring rights to the database entries of re3data.org.
Except where otherwise noted, content on this site is licensed under a Creative Commons Attribution 4.0 International License .
Cite this service: re3data.org - Registry of Research Data Repositories. https://doi.org/10.17616/R3D last accessed: 2024-04-19