Search | re3data.org

CLARIN.SI repository

Slovenian CLARIN repository

Subject(s)

Content type(s)

Country

CLARIN.SI is the Slovenian node of the European CLARIN (Common Language Resources and Technology Infrastructure) Centers. The CLARIN.SI repository is hosted at the Jožef Stefan Institute and offers long-term preservation of deposited linguistic resources, along with their descriptive metadata. The integration of the repository with the CLARIN infrastructure gives the deposited resources wide exposure, so that they can be known, used and further developed beyond the lifetime of the projects in which they were produced. Among the resources currently available in the CLARIN.SI repository are the multilingual MULTEXT-East resources, the CC version of Slovenian reference corpus Gigafida, the morphological lexicon Sloleks, the IMP corpora and lexicons of historical Slovenian, as well as many other resources for a variety of languages. Furthermore, several REST-based web services are provided for different corpus-linguistic and NLP tasks.

CLARIN-UK

Subject(s)

Content type(s)

Country

CLARIN-UK is a consortium of centres of expertise involved in research and resource creation involving digital language data and tools. The consortium includes the national library, and academic departments and university centres in linguistics, languages, literature and computer science.

Energy Data eXchange

EDX

Subject(s)

Content type(s)

Country

United States

The Energy Data eXchange (EDX) is an online collection of capabilities and resources that advance research and customize energy-related needs. EDX is developed and maintained by NETL-RIC researchers and technical computing teams to support private collaboration for ongoing research efforts, and tech transfer of finalized DOE NETL research products. EDX supports NETL-affiliated research by: Coordinating historical and current data and information from a wide variety of sources to facilitate access to research that crosscuts multiple NETL projects/programs; Providing external access to technical products and data published by NETL-affiliated research teams; Collaborating with a variety of organizations and institutions in a secure environment through EDX’s ;Collaborative Workspaces

Propylaeum@heiDATA

Data publications of the Specialized Information Service Classics

Subject(s)

Content type(s)

Country

Additional to the the e-publishing offer for articles, books and journals, Propylaeum provides classical scholars with the opportunity to archive the respective research data permanently. These can be linked directly to online publications hosted on the Heidelberg publishing platforms. All research data – e.g. images, videos, audio files, tables, graphics etc. – receive a DOI (Digital Object Identifiyer). Thus, they can be cited, viewed and permanently linked to as distinct academic output.

Yale-NUS Dataverse

Subject(s)

Content type(s)

Country

Singapore

Yale-NUS Dataverse is the institutional research data repository of Yale-NUS College. The goals of Yale-NUS Dataverse are to collect, preserve and showcase the research output of Yale-NUS researchers and through this, increase the research visibility of Yale-NUS researchers and demonstrate the research excellence of Yale-NUS College to the world.

SLAPIS

SLAPIS Early Warning System

Subject(s)

Content type(s)

Country

Italy
Niger

SLAPIS is an integrated Flood Early Warning System that aims to promote decision-making and behavioral changes from reactive to proactive at several levels, from the community to the administration, for the reduction of flood risk in the Communes of the Sirba (main tributary of the Niger River and cause of the main floods in the region)

CKAN @ IoT Lab

Subject(s)

Content type(s)

Country

IoT Lab is a research platform exploring the potential of crowdsourcing and Internet of Things for multidisciplinary research with more end-user interactions. IoT Lab is a European Research project which aims at researching the potential of crowdsourcing to extend IoT testbed infrastructure for multidisciplinary experiments with more end-user interactions. It addresses topics such as: - Crowdsourcing mechanisms and tools; - “Crowdsourcing-driven research”; - Virtualization of crowdsourcing and testbeds; - Ubiquitous Interconnection and Cloudification of testbeds; - Testbed as a Service platform; - Multidisciplinary experiments; - End-user and societal value creation; - Privacy and personal data protection.

INESC TEC Research Data Repository

Subject(s)

Content type(s)

Country

Portugal

The INESC TEC data repository showcases datasets produced or used by INESC TEC researchers and their partners. The repository is organized in four groups (institutional clusters). Computer Science, Power and Energy, Network and Intelligent Systems and Power and Energy.

GitHub

Subject(s)

Content type(s)

Country

United States

GitHub is the best place to share code with friends, co-workers, classmates, and complete strangers. Over three million people use GitHub to build amazing things together. With the collaborative features of GitHub.com, our desktop and mobile apps, and GitHub Enterprise, it has never been easier for individuals and teams to write better code, faster. Originally founded by Tom Preston-Werner, Chris Wanstrath, and PJ Hyett to simplify sharing code, GitHub has grown into the largest code host in the world.

ZFDM repository

FDR@UHH

Subject(s)

Content type(s)

Country

Germany

Research Data Repository of the Universität Hamburg

Angewandte Repository

Angewandte Repositorium

Subject(s)

Content type(s)

Country

Austria

Phaidra (Permanent Hosting, Archiving and Indexing of Digital Resources and Assets) is the University of Applied Arts Vienna System’s platform for long-term archiving of digital collections.

FULIR Data

Repozitorij istraživačkih podataka Instituta Ruđer Bošković

Subject(s)

Content type(s)

Standard office documents

Country

Croatia

FULIR Data is a research data repository that gathers, permanently stores and allows open access to primary data produced by researchers based at Ruđer Bošković Institute. Researchers deposit datasets by themselves (self-archiving) with the support given by the Centre for Scientific Information and their RDM experts.

UCL Research Data Repository

Subject(s)

Content type(s)

Country

United Kingdom

The UCL Research Data Repository is the institutional data repository of University College London. Based on Figshare, it accepts data deposits from UCL staff and doctoral students from all disciplines. Depositors are encouraged to use CC0 licences to make their data available to the widest possible range of users and the widest range of uses.

Forschungsdatenrepositorium der Otto-Friedrich-Universität Bamberg

Research Data Repository of the University of Bamberg

Subject(s)

Content type(s)

Country

Germany

Research Data Repository is the institutional repository of the University of Bamberg to enable storing, sharing and publishing of research data.

OpenML

Open Machine Learning

Subject(s)

Content type(s)

Country

OpenML is an open ecosystem for machine learning. By organizing all resources and results online, research becomes more efficient, useful and fun. OpenML is a platform to share detailed experimental results with the community at large and organize them for future reuse. Moreover, it will be directly integrated in today’s most popular data mining tools (for now: R, KNIME, RapidMiner and WEKA). Such an easy and free exchange of experiments has tremendous potential to speed up machine learning research, to engender larger, more detailed studies and to offer accurate advice to practitioners. Finally, it will also be a valuable resource for education in machine learning and data mining.

Loughborough Research Repository

Loughborough University Research Repository

Subject(s)

Content type(s)

Country

United Kingdom

Loughborough Research Repository is the institutional repository of Loughborough University powered by figshare.

CLARIN repository at the University of Tübingen

CLARIN Center Tübingen

Subject(s)

Content type(s)

Country

The repository is part of the National Research Data Infrastructure initiative Text+, in which the University of Tübingen is a partner. It is housed at the Department of General and Computational Linguistics. The infrastructure is maintained in close cooperation with the Digital Humanities Centre, which is a core facility of the university, colaborating with the library and computing center of the university. Integration of the repository into the national CLARIN-D and international CLARIN infrastructures gives it wide exposure, increasing the likelihood that the resources will be used and further developed beyond the lifetime of the projects in which they were developed. Among the resources currently available in the Tübingen Center Repository, researchers can find widely used treebanks of German (e.g. TüBa-D/Z), the German wordnet (GermaNet), the first manually annotated digital treebank (Index Thomisticus), as well as descriptions of the tools used by the WebLicht ecosystem for natural language processing.

CESA | Repositorio de datos académicos

Subject(s)

Content type(s)

Scientific and statistical data formats

Country

Colombia

Discover the data on entrepreneurship projects, innovation plans, digital transformation proposals, consumers, and financial markets. Also, explore research on business, management, and entrepreneurship research development at our Business school.

CITK

Cognitive Interaction Toolkit

Subject(s)

Content type(s)

Country

Germany

<<<!!!<<< This repository is no longer available. >>>!!!>>>

Stockholm University Library Dataverse

Subject(s)

Content type(s)

Country

For datasets from individual researchers or research groups affiliated with Stockholm University, who do not want set up a separate Dataverse for a project or institution. Metadata provisions for Geospatial, Social Science, Humanities, Astronomy, Astrophysics, Life Sciences and Journals (all optional, by choice) are included. Data curation help from Stockholm University Library possible on request.

NIST Data Discovery

NIST PDR

Subject(s)

Content type(s)

Country

United States

Data products developed and distributed by the National Institute of Standards and Technology span multiple disciplines of research and are widely used in research and development programs by industry and academia. NIST's publicly available data sets showcase its committment to providing accurate, well-curated measurements of physical properties, exemplified by the Standard Reference Data program, as well as its committment to advancing basic research. In accordance with U.S. Government Open Data Policy and the NIST Plan for providing public access to the results of federally funded research data, NIST maintains a publicly accessible listing of available data, the NIST Public Dataset List (json). Additionally, these data are assigned a Digital Object Identifier (DOI) to increase the discovery and access to research output; these DOIs are registered with DataCite and provide globally unique persistent identifiers. The NIST Science Data Portal provides a user-friendly discovery and exploration tool for publically available datasets at NIST. This portal is designed and developed with data.gov Project Open Data standards and principles. The portal software is hosted in the usnistgov github repository.

CLAPOP

The Dutch CLARIN Portal Pages

Subject(s)

Content type(s)

Country

CLAPOP is the portal of the Dutch CLARIN community. It brings together all relevant resources that were created within the CLARIN NL project and that now are part of the CLARIN NL infrastructure or that were created by other projects but are essential for the functioning of the CLARIN (NL) infrastructure. CLARIN-NL has closely cooperated with CLARIN Flanders in a number of projects. The common results of this cooperation and the results of this cooperation created by CLARIN Flanders are included here as well.

KONECT

Koblenz Network Collection

Subject(s)

Content type(s)

Country

This is the KONECT project, a project in the area of network science with the goal to collect network datasets, analyse them, and make available all analyses online. KONECT stands for Koblenz Network Collection, as the project has roots at the University of Koblenz–Landau in Germany. All source code is made available as Free Software, and includes a network analysis toolbox for GNU Octave, a network extraction library, as well as code to generate these web pages, including all statistics and plots. KONECT contains over a hundred network datasets of various types, including directed, undirected, bipartite, weighted, unweighted, signed and rating networks. The networks of KONECT are collected from many diverse areas such as social networks, hyperlink networks, authorship networks, physical networks, interaction networks and communication networks. The KONECT project has developed network analysis tools which are used to compute network statistics, to draw plots and to implement various link prediction algorithms. The result of these analyses are presented on these pages. Whenever we are allowed to do so, we provide a download of the networks.

University of Guelph Research Data Repositories

University of Guelph Dataverse

Subject(s)

Content type(s)

Country

The University of Guelph Research Data Repositories provide long-term stewardship of research data created at or in cooperation with the University of Guelph. The Data Repositories are guided by the FAIR Guiding Principles for scientific data management and stewardship which aim to improve the Findability, Accessibility, Interoperability and Reuse of research data. The Data Repositories is composed of two main collections: the Agri-environmental Research Data collection which houses agricultural and environmental research data, and the Cross-disciplinary Research Data collection which houses all other disciplinary research data.

SinMin

Subject(s)

Content type(s)

Country

United States

Sinmin contains texts of different genres and styles of the modern and old Sinhala language. The main sources of electronic copies of texts for the corpus are online Sinhala newspapers, online Sinhala news sites, Sinhala school textbooks available in online, online Sinhala magazines, Sinhala Wikipedia, Sinhala fictions available in online, Mahawansa, Sinhala Blogs, Sinhala subtitles and Sri lankan gazette.

Subjects

Content Types

Countries

AID systems

API

Certificates

Data access

Data access restrictions

Database access

Database licenses

Data licenses

Data upload

Data upload restrictions

Enhanced publication

Institution responsibility type

Institution type

Keywords

Metadata standards

PID systems

Provider types

Quality management

Repository languages

Software

Syndications

Repository types

Versioning