Polish RoBERTA model trained on Polish literature, Wikipedia, and Oscar. The major assumption is that quality text will give a good model.
☆34May 25, 2021Updated 5 years ago
Alternatives and similar repositories for PoLitBert
Users that are interested in PoLitBert are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- RoBERTa models for Polish☆90Mar 8, 2022Updated 4 years ago
- Evaluation of Sentence Representations in Polish☆22Dec 29, 2022Updated 3 years ago
- A curated list of resources dedicated to Natural Language Processing (NLP) in polish. Models, tools, datasets.☆307Aug 8, 2021Updated 5 years ago
- ☆51Aug 22, 2022Updated 3 years ago
- Generic framework for information extraction tasks, including recognition of named entities, temporal expressions, spatial expressions an…☆13Jun 5, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Polish BERT☆72Oct 27, 2020Updated 5 years ago
- ☆14Mar 28, 2025Updated last year
- Code and data accompanying the paper "Approaching nested named entity recognition with parallel LSTM-CRFs."☆27Dec 8, 2022Updated 3 years ago
- Resources for doing NLP in Polish☆48Nov 4, 2019Updated 6 years ago
- HerBERT is a BERT-based Language Model trained on Polish Corpora using only MLM objective with dynamic masking of whole words.☆76Feb 3, 2022Updated 4 years ago
- ☆35Updated this week
- Polish datsets for grammatical error correction☆12Oct 13, 2023Updated 2 years ago
- Neuro-symbolic concept embedding and reasoning for ALC knowledge bases☆14Feb 7, 2023Updated 3 years ago
- Fine-tuning scripts for evaluating transformer-based models on KLEJ benchmark.☆29Jul 6, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- COMBO is jointly trained tagger, lemmatizer and dependency parser.☆36Mar 24, 2023Updated 3 years ago
- A morphosyntactic tagger for Polish based on conditional random fields☆23Apr 6, 2021Updated 5 years ago
- ☆13Aug 13, 2020Updated 5 years ago
- PolEval 2021 Task 1☆15Jun 28, 2022Updated 4 years ago
- Solutions to the labs and exercises in ISL.☆11Jan 21, 2019Updated 7 years ago
- ☆16Jul 10, 2023Updated 3 years ago
- Tool for named entity recognition for Polish based on deep learning.☆32Mar 24, 2023Updated 3 years ago
- Source code for "Taming GANs with Lookahead–Minmax", ICLR 2021.☆15Mar 28, 2021Updated 5 years ago
- MediaWiki Categories Model☆13Feb 14, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official code for "Advancing Multimodal Large Language Models with Quantization-Aware Scale Learning for Efficient Adaptation" | [MM2…☆14Dec 7, 2024Updated last year
- IPython magic for simple, organized, compressed and encrypted: storage & transfer of files between notebooks.☆13Apr 13, 2026Updated 3 months ago
- Example for Logging LLM Evaluator Prompt Responses☆18Aug 14, 2023Updated 2 years ago
- Representation Learning of Entities and Documents from Knowledge Base Descriptions☆18Oct 6, 2018Updated 7 years ago
- A Node.js tool to examine the correctness of Open Data Metadata and build custom dataset profiles☆12Sep 26, 2023Updated 2 years ago
- ☆14Aug 28, 2019Updated 6 years ago
- a detail tutorials of allennlp , which is based on my own view.☆10Mar 7, 2020Updated 6 years ago
- NLP Model for predicting 17 different languages☆16Oct 19, 2023Updated 2 years ago
- A simple Panel-based dashboard visualizing geotagged tweets with hvplot and Datashader.☆16Mar 25, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Using synthetically generated video data to learn how/when R-CNNs can outperform CNNs.☆17Mar 18, 2018Updated 8 years ago
- Java Speech Toolkit☆12Apr 25, 2021Updated 5 years ago
- communication sur le moteur de pseudonymisation de la Cour de Cassation☆21Feb 14, 2023Updated 3 years ago
- Survey of available speech datasets for Polish ASR development☆17Jan 1, 2025Updated last year
- A literature review for constructing and using knowledge graphs in a biomedical setting.☆11May 22, 2020Updated 6 years ago
- Code from the paper "Specializing Unsupervised Pretraining Models for Word-Level Semantic Similarity"☆19May 8, 2020Updated 6 years ago
- ☆10Jul 18, 2018Updated 8 years ago