German Dataset for Legal Information Retrieval
☆27Feb 26, 2024Updated 2 years ago
Alternatives and similar repositories for GerDaLIR
Users that are interested in GerDaLIR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A dataset of semantically related sentence pairs in the German legal domain☆10Feb 26, 2021Updated 5 years ago
- Enhancing Legal Case Retrieval via Scaling High-quality Synthetic Query-Candidate Pairs (EMNLP 2024)☆18Nov 17, 2024Updated last year
- A collection of datasets and other resources for legal text processing.☆301Aug 24, 2026Updated last week
- ☆17Jun 6, 2024Updated 2 years ago
- This is a german text corpus from Wikipedia. It is cleaned, preprocessed and sentence splitted. It's purpose is to train NLP embeddings l…☆24Feb 22, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PyTorch code for JEREX: Joint Entity-Level Relation Extractor☆68Dec 9, 2021Updated 4 years ago
- A Dataset of German Legal Documents for Named Entity Recognition☆179Oct 19, 2022Updated 3 years ago
- Regulärer Ausdruck zum Finden von Gesetzen in Texten/Regex to find German laws.☆21Jul 18, 2023Updated 3 years ago
- Python code to automatically produce a summary of a piece of text.☆11Sep 8, 2016Updated 9 years ago
- Code repository of the NAACL'21 paper "CoRT: Complementary Rankings from Transformers"☆12Jul 7, 2021Updated 5 years ago
- Ukrainian ELECTRA model☆12Mar 11, 2023Updated 3 years ago
- The Python Digital Toolbox contains examples of how to solve various data analysis problems using Python libraries.☆16Aug 24, 2026Updated last week
- Convert legal statutes and cases from official sources (or juris) to graphs☆32Sep 11, 2025Updated 11 months ago
- ☆14Jun 2, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- German Parliamentary Corpus (GerParCor)☆32Mar 29, 2026Updated 5 months ago
- A daily archive of https://www.gesetze-im-internet.de☆42Updated this week
- ☆15Nov 14, 2022Updated 3 years ago
- A rolling version of the Latent Dirichlet Allocation.☆13Updated this week
- Submissions, baselines and evaluations scripts for the 2nd version of the WebNLG+ Challenge 2020☆13Feb 1, 2022Updated 4 years ago
- Code for the paper "A Comprehensive Evaluation of Large Language Models on Legal Judgment Prediction"☆13Oct 20, 2023Updated 2 years ago
- ☆13Oct 28, 2024Updated last year
- Poetry Corpora Annotated on Aesthetic Emotions☆12Aug 2, 2022Updated 4 years ago
- A really fast document ranking engine using BM25 and TF-IDF. Based on Python using NLP packages NLTK and spacY.☆17May 8, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- TextComplexityDE dataset consists of 1000 sentences in the German language with subjective complexity rating, collected from German learn…☆12Apr 8, 2022Updated 4 years ago
- A DH abstracts conversion tool☆13Apr 24, 2026Updated 4 months ago
- Source code for EMNLP 2023 paper "Probabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex Questions".☆23Mar 21, 2024Updated 2 years ago
- Implements SemRe-Rank: improving automatic term extraction by incorporating semantic relatedness with personalised pagerank☆16Apr 7, 2018Updated 8 years ago
- EQUATE (Evaluating Quantitative Understanding Aptitude in Textual Entailment), framework for evaluating quantitative reasoning ability in…☆14Feb 13, 2022Updated 4 years ago
- Mining Legal Arguments in Court Decisions - Data and software☆81May 15, 2023Updated 3 years ago
- Annotated data set consisting of user comments posted to a German-language newspaper website☆18Jun 28, 2018Updated 8 years ago
- A curated list of awesome resources to create and customize your Curriculum Vitae☆33Updated this week
- code for our EMNLP2020 paper: Multilevel Text Alignment with Cross-Document Attention by Xuhui Zhou, Nikolaos Pappas, and Noah A. Smith☆14May 18, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 📜 Dehyphenation of broken text (mainly German), i.e., extracted from a PDF☆39Mar 8, 2022Updated 4 years ago
- TREC QA dataset for question answering cleaned for usage in Question Answering☆14Aug 26, 2019Updated 7 years ago
- Source code for the AI2 Reasoning Challenge (ARC) submission.☆16Dec 8, 2022Updated 3 years ago
- The dataset consists of public social media url pairs and the corresponding entailment label for an external conference (ACL 2021). Each …☆14Aug 16, 2021Updated 5 years ago
- Fork of RecurrentGPT with modifications☆10Sep 18, 2024Updated last year
- ☆16Oct 6, 2022Updated 3 years ago
- HiCAL is a system for efficient high-recall retrieval with an adaptable assessing interface.☆38Dec 26, 2022Updated 3 years ago