Search | re3data.org

Filter

Reset all

Subjects

Content Types

Countries

AID systems

API

Certificates

Data access

Data access restrictions

Database access

Database access restrictions

Database licenses

Data licenses

Data upload

Data upload restrictions

Enhanced publication

unknown (122)

Institution responsibility type

Institution type

Keywords

Metadata standards

PID systems

Provider types

Quality management

Repository languages

Software

Syndications

Repository types

Versioning

Toogle short help

* at the end of a keyword allows wildcard searches
" quotes can be used for searching phrases
+ represents an AND search (default)
| represents an OR search
- represents a NOT operation
( and ) implies priority
~N after a word specifies the desired edit distance (fuzziness)
~N after a phrase specifies the desired slop amount

← Previous
1 (current)
2
3
4
5
Next →

Found 122 result(s)

Deutsches Textarchiv

DTA

Subject(s)

Content type(s)

Country

The German Text Archive (Deutsches Textarchiv, DTA) presents online a selection of key German-language works in various disciplines from the 17th to 19th centuries. The electronic full-texts are indexed linguistically and the search facilities tolerate a range of spelling variants. The DTA presents German-language printed works from around 1650 to 1900 as full text and as digital facsimile. The selection of texts was made on the basis of lexicographical criteria and includes scientific or scholarly texts, texts from everyday life, and literary works. The digitalisation was made from the first edition of each work. Using the digital images of these editions, the text was first typed up manually twice (‘double keying’). To represent the structure of the text, the electronic full-text was encoded in conformity with the XML standard TEI P5. The next stages complete the linguistic analysis, i.e. the text is tokenised, lemmatised, and the parts of speech are annotated. The DTA thus presents a linguistically analysed, historical full-text corpus, available for a range of questions in corpus linguistics. Thanks to the interdisciplinary nature of the DTA Corpus, it also offers valuable source-texts for neighbouring disciplines in the humanities, and for scientists, legal scholars and economists.

Repository CLARIN Centre Leipzig

CLARIN repository at the Saxon Academy of Sciences and Humanities

Subject(s)

Content type(s)

Country

The CLARIN/Text+ repository at the Saxon Academy of Sciences and Humanities in Leipzig offers longterm preservation of digital resources, along with their descriptive metadata. The mission of the repository is to ensure the availability and longterm preservation of resources, to preserve knowledge gained in research, to aid the transfer of knowledge into new contexts, and to integrate new methods and resources into university curricula. Among the resources currently available in the Leipzig repository are a set of corpora of the Leipzig Corpora Collection (LCC), based on newspaper, Wikipedia and Web text. Furthermore several REST-based webservices are provided for a variety of different NLP-relevant tasks The repository is part of the CLARIN infrastructure and part of the NFDI consortium Text+. It is operated by the Saxon Academy of Sciences and Humanities in Leipzig.

Kujawsko-Pomorska Digital Library

Kujawsko-Pomorska Biblioteka Cyfrowa

Subject(s)

Content type(s)

Country

Poland

The KPDL covers cultural heritage, scientific and regional collections – digital copies of different forms of publications: books, journals, graphics, articles, leaflets, posters, playbills, photographs, invitations, maps, exhibition catalogues and trade fairs of the region. The Kujawsko-Pomorska Digital Library is to serve scientists, students, schoolchildren and all the citizens of the region.

ScholarWorks Boise State University

Subject(s)

Content type(s)

Country

United States

ScholarWorks is a collection of services designed to capture and showcase all scholarly output by the Boise State University community.

CLARIN service center of the Zentrum Sprache at the BBAW

CLARIN Center BBAW

Subject(s)

Content type(s)

Structured text

Country

The Berlin-Brandenburg Academy of Sciences and Humanities (BBAW) is a CLARIN partner institution and has been an officially certified CLARIN service center since June 20th, 2013. The CLARIN center at the BBAW focuses on historical text corpora (predominantly provided by the 'Deutsches Textarchiv'/German Text Archive, DTA) as well as on lexical resources (e.g. dictionaries provided by the 'Digitales Wörterbuch der Deutschen Sprache'/Digital Dictionary of the German Language, DWDS).

CLARIND-UdS: Language Resources Repository at UdS

CLARIND-UdS, Repositorium für Sprachressourcen an der Universität des Saarlandes

Subject(s)

Content type(s)

Country

In collaboration with other centres in the Text+ consortium and in the CLARIN infrastructure, the CLARIND-UdS enables eHumanities by providing a service for hosting and processing language resources (notably corpora) for members of the research community. CLARIND-UdS centre thus contributes of lifting the fragmentation of language resources by assisting members of the research community in preparing language materials in such a way that easy discovery is ensured, interchange is facilitated and preservation is enabled by enriching such materials with meta-information, transforming them into sustainable formats and hosting them. We have an explicit mission to archive language resources especially multilingual corpora (parallel, comparable) and corpora including specific registers, both collected by associated researchers as well as researchers who are not affiliated with us.

PolMine

PolMine Project

Subject(s)

Content type(s)

Country

The focus of PolMine is on texts published by public institutions in Germany. Corpora of parliamentary protocols are at the heart of the project: Parliamentary proceedings are available for long stretches of time, cover a broad set of public policies and are in the public domain, making them a valuable text resource for political science. The project develops repositories of textual data in a sustainable fashion to suit the research needs of political science. Concerning data, the focus is on converting text issued by public institutions into a sustainable digital format (TEI/XML).

Kielipankki

The Language Bank of Finland

Subject(s)

Content type(s)

Country

The Language Bank features text and speech corpora with different kinds of annotations in over 60 languages. There is also a selection of tools for working with them, from linguistic analyzers to programming environments. Corpora are also available via web interfaces, and users can be allowed to download some of them. The IP holders can monitor the use of their resources and view user statistics.

SEDICI

Institutional Repository of the UNLP

Subject(s)

Content type(s)

Country

Argentina

SEDICI [Intellectual Creation Diffusion Service] is the institutional repository of the Universidad Nacional de La Plata (UNLP), a public university located at Argentina. Its goal is to index, preserve and grant open access to all kind of academic work produced at this institution including thesis, scientific articles, datasets, books, conference objects, and more.

RosDok

RosDok - Rostocker Dokumentenserver

Subject(s)

Content type(s)

Country

Germany

RosDok is the platform of the University of Rostock for online publication and permanent archiving of digital documents.

CSIRO Data Access Portal

CSIRO DAP

Subject(s)

Content type(s)

Country

Australia

The CSIRO Data Access Portal provides access to data published by CSIRO across a range of disciplines to facilitate sharing and reuse of data held by CSIRO.

Das Deutsche Referenzkorpus

DeReKo

Subject(s)

Content type(s)

Country

The project is set up in order to improve the infrastructure for text-based linguistic research and development by building a huge, automatically annotated German text corpus and the corresponding tools for corpus annotation and exploitation. DeReKo constitutes the largest linguistically motivated collection of contemporary German texts, contains fictional, scientific and newspaper texts, as well as several other text types, contains only licenced texts, is encoded with rich meta-textual information, is fully annotated morphosyntactically (three concurrent annotations), is continually expanded, with a focus on size and stratification of data, may be analyzed free of charge via the query system COSMAS II, serves as a 'primordial sample' from which users may draw specialized sub-samples (socalled 'virtual corpora') to represent the language domain they wish to investigate. !!! Access to data of Das Deutsche Referenzkorpus is also provided by: IDS Repository https://www.re3data.org/repository/r3d100010382 !!!

CLARIN repository at the University of Tübingen

CLARIN Center Tübingen

Subject(s)

Content type(s)

Country

The repository is part of the National Research Data Infrastructure initiative Text+, in which the University of Tübingen is a partner. It is housed at the Department of General and Computational Linguistics. The infrastructure is maintained in close cooperation with the Digital Humanities Centre, which is a core facility of the university, colaborating with the library and computing center of the university. Integration of the repository into the national CLARIN-D and international CLARIN infrastructures gives it wide exposure, increasing the likelihood that the resources will be used and further developed beyond the lifetime of the projects in which they were developed. Among the resources currently available in the Tübingen Center Repository, researchers can find widely used treebanks of German (e.g. TüBa-D/Z), the German wordnet (GermaNet), the first manually annotated digital treebank (Index Thomisticus), as well as descriptions of the tools used by the WebLicht ecosystem for natural language processing.

Portal Zarządzania Wiedzą UJ CM

Knowledge Management Platform of Jagiellonian University Medical College

Subject(s)

Content type(s)

Country

Poland

Portal Zarządzania Wiedzą UJ CM is a knowledge and research potential management platform of medical information. Easy data localization will be possible thanks to DOI and URL addresses given by administrators of the Knowledge Management Platform JU MC. The data will be stored in 2 copies for 10 years.

NAKALA

Subject(s)

Humanities and Social Sciences

Content type(s)

Country

France

NAKALA is a repository dedicated to SSH research data in France. Given its generalist and multi-disciplinary nature, all types of data are accepted, although certain formats are recommended to ensure longterm data preservation. It has been developed and is hosted by Huma-Num, the French national research infrastructure for digital humanities.

data_UMR

Subject(s)

Content type(s)

Country

Germany

data_UMR is the institutional repository of the University of Marburg for research data of all kinds that has been generated in the context of research activities at the University of Marburg.

CiteSeerX

Subject(s)

Content type(s)

Country

United States

CiteSeerx is an evolving scientific literature digital library and search engine that focuses primarily on the literature in computer and information science. CiteSeerx aims to improve the dissemination of scientific literature and to provide improvements in functionality, usability, availability, cost, comprehensiveness, efficiency, and timeliness in the access of scientific and scholarly knowledge. Rather than creating just another digital library, CiteSeerx attempts to provide resources such as algorithms, data, metadata, services, techniques, and software that can be used to promote other digital libraries. CiteSeerx has developed new methods and algorithms to index PostScript and PDF research articles on the Web.

Australian SuperSite Network Data Portal

SupeSites Data Portal

Subject(s)

Content type(s)

Country

Australia

<<<!!!<<< This repository is no longer available. It is part now of TERN Data Discovery Portal https://www.re3data.org/repository/r3d100012013 >>>!!!>>>

mdw Repository

Subject(s)

Content type(s)

Country

Austria

mdw Repository provides researchers with a robust infrastructure for research data management and ensures accessibility of research data during and after completion of research projects, thus, providing a quality boost to contemporary and future research.

RODARE

Rossendorf Data Repository

Subject(s)

Content type(s)

Country

Germany

Rodare is the institutional research data repository at HZDR (Helmholtz-Zentrum Dresden-Rossendorf). Rodare allows HZDR researchers to upload their research software and data and enrich those with metadata to make them findable, accessible, interoperable and retrievable (FAIR). By publishing all associated research software and data via Rodare research reproducibility can be improved. Uploads receive a Digital Object Identfier (DOI) and can be harvested via a OAI-PMH interface.

B2SHARE Server Forschungszentrum Jülich

Subject(s)

Content type(s)

Country

Germany

B2SHARE allows publishing research data and belonging metadata. It supports different research communities with specific metadata schemas. This server is provided for researchers of the Research Centre Juelich and related communities.

Research Data Repository FDAT

Research Data Repository of the University of Tuebingen

Subject(s)

Content type(s)

Country

Germany

FDAT is a research data repository hosted by the University of Tübingen, designed to facilitate long-term archiving and publication of research data. Managed by the Information, Communication and Media Center (IKM), it primarily caters to the humanities and social sciences, while welcoming researchers from all scientific disciplines at the university. Committed to high-quality data management, FDAT emphasizes the importance of adhering to the FAIR Data Principles, promoting findability, accessibility, interoperability, and reusability of the research data it contains.

OLOS

Subject(s)

Content type(s)

Country

Switzerland

OLOS is a Swiss-based data management portal tailored for researchers and institutions. Powerful yet easy to use, OLOS works with most tools and formats across all scientific disciplines to help researchers safely manage, publish and preserve their data. The solution was developed as part of a larger project focusing on Data Life Cycle Management (dlcm.ch) that aims to develop various services for research data management. Thanks to its highly modular architecture, OLOS can be adapted both to small institutions that need a "turnkey" solution and to larger ones that can rely on OLOS to complement what they have already implemented. OLOS is compatible with all formats in use in the different scientific disciplines and is based on modern technology that interconnects with researchers' environments (such as Electronic Laboratory Notebooks or Laboratory Information Management Systems).

eData: the STFC Research Data Repository

Subject(s)

Content type(s)

Country

United Kingdom

eData is an institutional repository where STFC staff can deposit data and software that underpin journal articles and other published research.

LINDAT/CLARIAH-CZ repository

LINDAT/CLARIN repository

Subject(s)

Content type(s)

Country

LINDAT/CLARIN is designed as a Czech “node” of Clarin ERIC (Common Language Resources and Technology Infrastructure). It also supports the goals of the META-NET language technology network. Both networks aim at collection, annotation, development and free sharing of language data and basic technologies between institutions and individuals both in science and in all types of research. The Clarin ERIC infrastructural project is more focused on humanities, while META-NET aims at the development of language technologies and applications. The data stored in the repository are already being used in scientific publications in the Czech Republic. In 2019 LINDAT/CLARIAH-CZ was established as a unification of two research infrastructures, LINDAT/CLARIN and DARIAH-CZ.

← Previous
1 (current)
2
3
4
5
Next →

Current projects
EOSC FAIR-IMPACT

re3data COREF

To the extent possible under law, re3data.org has waived all copyright and related or neighboring rights to the database entries of re3data.org.
Except where otherwise noted, content on this site is licensed under a Creative Commons Attribution 4.0 International License .
Cite this service: re3data.org - Registry of Research Data Repositories. https://doi.org/10.17616/R3D last accessed: 2024-06-19