Factsheet
Rezension von „InterCorp – Ein mehrsprachiges Parallelkorpus des Tschechischen Nationalkorpus“ (Český národní korpus)
Reviewed resource
People
Editor
- Alexandr Rosen
- Martin Vavřín
- Adrian Zasina
Questionnaire
General information
Can the text collection be identified in terms similar to traditional bibliographic descriptions (title, responsible editors, institution, date(s) of publication, identifier/address)? cf. Catalogue 1.1
Answer Yes
Are the contributors (editors, institutions, associates) of the project documented? cf. Catalogue 1.3
Answer Yes
Is contact information given? cf. Catalogue 1.4
Answer Yes
Aims
Is there a description of the aims and contents of the text collection? cf. Catalogue 2.1
Answer Yes
What is the purpose of the text collection? cf. Catalogue 2.2
Answer research, teaching, purpose, (option)
What kind of research does the collection allow to conduct primarily? cf. Catalogue 3.1.8
Answer qualitative
How does the text collection classify itself (e.g. in its title or documentation)? cf. Catalogue 2.3
Answer corpus
To which field(s) of research does the text collection contribute? cf. Catalogue 2.2
Answer linguistics, (option)
Content
What era(s) do the texts belong to? cf. Catalogue 2.5
Answer contemporary
What languages are the texts in? cf. Catalogue 2.5
Answer arabic, chinese, danish, english, finnish, french, german, greek, hebrew, hindi, italian, japanese, norwegian, polish, portuguese, russian, spanish, swedish, turkish, (option)
What kind of texts are in the collection? cf. Catalogue 2.5
Answer works, articles, charters, protocols, (option)
What kind of information is published in addition to the texts? cf. Catalogue 2.5
Answer (option)
Composition
Are the principles and decisions regarding the design of the text collection, its composition and the selection of texts documented? cf. Catalogue 3.1.1-3.1.3
Answer Yes
What selection criteria have been chosen for the text collection? cf. Catalogue 3.1
Answer language, epoch, (option)
How large is the text collection in number of texts/records? cf. Catalogue 3.1.4
Answer gt1000
How large is the text collection in number of tokens? cf. Catalogue 3.1.4
Answer gt10mio
Does the text collection have identifiable sub-collections or components? cf. Catalogue 3.1.5
Answer Yes
acquisition
Answer Not answered
Has the quality of the data (transcriptions, metadata, annotations, etc.) been checked? cf. Catalogue 3.1.7
Answer Yes
Considering aims and methods of the text collection, how would you classify it further? For definitions please consider the help-texts. cf. Catalogue 3.1.8
Answer corpus, corpus
Data modelling
How are the textual sources represented in the digital collection? cf. Catalogue 3.2.1
Answer (option)
In which basic format are the texts encoded? cf. Catalogue 3.2.4
Answer xml
With what information are the texts further enriched? cf. Catalogue 3.2.2
Answer annotations, information
How are the annotations linked to the texts themselves? cf. Catalogue 3.2.2
Answer embedded
What kind of metadata are included in the text collection? cf. Catalogue 3.2.3
Answer descriptive, structural, administrative
On which level are the metadata included? cf. Catalogue 3.2.2
Answer parts, texts
What kind of data/metadata/annotation schemas are used for the text collection? cf. Catalogue 3.2.4
Answer schema, schema
Which standards for text encoding, metadata and annotation are used in the text collection? cf. Catalogue 3.2.4
Answer tagsets
Provision
Is the textual data accessible in a source format (e.g. XML, TXT)? cf. Catalogue 4.1
Answer No
Can the entire raw data of the project be downloaded (as a whole)? cf. Catalogue 4.2
Answer No
Are there technical interfaces which allow the reuse of the data of the text collection in other contexts? cf. Catalogue 4.2
Answer (option)
Besides the textual data, does the project provide analytical data (e.g. statistics) to download or harvest? cf. Catalogue 4.3
Answer No
Can you use the data with other tools useful for this kind of content? cf. Catalogue 4.4
Answer Yes
User interface
Does the text collection have a dedicated user interface designed for the collection at hand in which the texts of the collection are represented and/or in which the data is analyzable? cf. Catalogue 5.1
Answer Yes
user_interface_sub
Answer From your point of view, is the interface of the text collection clearly arranged and easy to navigate so that the user can quickly identify the purpose, the content and the main access methods of the resource?, Is there a personalisation mode that enables the users e.g. to create their own sub-collections of the existing text collection?
Preservation
Does the text collection provide sufficient documentation about the project in general as well as about the aims, contents and methods of the text collection? cf. Catalogue 6.1
Answer Yes
Is the text collection Open Access? cf. Catalogue 6.2
Answer Yes
rights
Answer Are the rights to (re)use the content declared?
Are there persistent identifiers and an addressing system for the text collection and/or parts/objects of it and which mechanism is used to that end? cf. Catalogue 6.3
Answer (option)
Does the text collection supply citation guidelines? cf. Catalogue 6.3
Answer Yes
Does the documentation include information about the long term sustainability of the basic data (archiving of the data)? cf. Catalogue 6.4
Answer No
Does the project provide information about institutional support for the curation and sustainability of the project? cf. Catalogue 6.4
Answer Yes
Is the text collection completed? cf. Catalogue 6.4
Answer Not evaluated