This is a new version of the repository. Do let us know (lindat-help at ufal.mff.cuni.cz) if you encounter any issues.

Annotated Corpus of Czech Case Law for Reference Recognition Tasks

Please use the following text to cite this item or export to a predefined format:
Harašta, Jakub; et al., 2018, Annotated Corpus of Czech Case Law for Reference Recognition Tasks, LINDAT/CLARIAH-CZ digital library at the Institute of Formal and Applied Linguistics (ÚFAL), http://hdl.handle.net/11234/1-2647.
Date issued
2018
Language(s)
Description
Annotated corpus of 350 decision of Czech top-tier courts (Supreme Court, Supreme Administrative Court, Constitutional Court). Every decision is annotated by two trained annotators and then manually adjudicated by one trained curator to solve possible disagreements between annotators. Adjudication was conducted non-destructively, therefore dataset contains all original annotations. Corpus was developed as training and testing material for reference recognition tasks. Dataset contains references to other court decisions and literature. All references consist of basic units (identifier of court decision, identification of court issuing referred decision, author of book or article, title of book or article, point of interest in referred document etc.), values (polarity, depth of discussion etc.).
Acknowledgement
This item isPublicly Available
and licensed under:

Files in this item

Name
corpus.json
Size
30.97 MB
Format
application/octet-stream
Description
Corpus
MD5
755c97e9e6f272651faab9cfae5117c2
Preview
  File Preview
Name
ReadMe.zip
Size
224.25 KB
Format
application/zip
Description
ReadMe
MD5
58c38afec21a555b97c1de57c2d1e8a0
Preview
  File Preview
    • README.md5 kB
    • FigPipeline.png81 kB
    • FigScheme.png165 kB