Train punctuation and capitalization models for different languages
☆26Apr 2, 2022Updated 4 years ago
Alternatives and similar repositories for multipunct
Users that are interested in multipunct are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Convert MUSE from TensorFlow to PyTorch and ONNX☆11May 22, 2024Updated 2 years ago
- Репозиторий измеряет качество Yandexgpt, Gigachat, T-Pro, Saiga, Vikhr, Ruadapt на популярных англоязычных бенчмарках: MGSM, MATH, HumanE…☆25Apr 16, 2025Updated last year
- RuLeanALBERT is a pretrained masked language model for the Russian language that uses a memory-efficient architecture.☆92May 27, 2023Updated 3 years ago
- 🇷🇺 Punctuation restoration production-ready model for Russian language 🇷🇺☆59Jul 9, 2021Updated 5 years ago
- ☆13Aug 7, 2021Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A database-like benchmark of feature generation from time-series data☆13Nov 27, 2024Updated last year
- Проект для перевода чисел, записанных в текстовом виде на русском языке.☆11Apr 5, 2022Updated 4 years ago
- ☆22Jun 11, 2026Updated last month
- Russian text normalization pipeline for speech-to-text and other applications based on tagging s2s networks☆122Mar 15, 2021Updated 5 years ago
- Reinforcement Learning Library.☆29Aug 16, 2022Updated 3 years ago
- AWD-LSTM language model trained on newspaper corpora with fast.ai☆27Apr 9, 2020Updated 6 years ago
- This repository provides data and code for "Vox Populi, Vox DIY: Benchmark Dataset for Crowdsourced Audio Transcription" paper.☆16Jul 22, 2021Updated 4 years ago
- Git Hooks Tutorial.☆17Jul 6, 2022Updated 4 years ago
- ☆13Dec 7, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The broad index of NLP resources for Eastern European languages. The best EEML 2021 project.☆19Jun 24, 2022Updated 4 years ago
- CraftML is a restful web service for easy pipeline creation without code.☆13Apr 18, 2021Updated 5 years ago
- ☆17Apr 14, 2023Updated 3 years ago
- Pipeline for training NER models using PyTorch.☆55Jul 19, 2022Updated 4 years ago
- SAGE: Spelling correction, corruption and evaluation for multiple languages☆167Dec 8, 2025Updated 7 months ago
- ☆23Aug 26, 2024Updated last year
- Top ML papers of the week.☆46Updated this week
- Unofficial implementation of QaNER: Prompting Question Answering Models for Few-shot Named Entity Recognition.☆63Oct 15, 2022Updated 3 years ago
- ☆15Sep 15, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- EMNLP 2024 | Style-Specific Neurons for Steering LLMs in Text Style Transfer☆14Mar 23, 2025Updated last year
- Make GNN easy to start with☆134Mar 10, 2026Updated 4 months ago
- Repository containing our datasets for HTR (handwritten text recognition) task.☆27Sep 8, 2022Updated 3 years ago
- Code for "ParaGuide: Guided Diffusion Paraphrasers for Plug-and-Play Textual Style Transfer"☆16Jul 17, 2024Updated 2 years ago
- Question answering on russian with XLMRobertaLarge as a service☆21Nov 6, 2021Updated 4 years ago
- ML Course created for Bauman Moscow State Technical University☆66Aug 31, 2022Updated 3 years ago
- Code and data for the NAACL 2021 paper: "XFORMAL: A Benchmark for Multilingual Formality Style Transfer"☆12Jun 7, 2021Updated 5 years ago
- Russian coreference resolution made as simple and accessible as could be☆11Sep 3, 2022Updated 3 years ago
- Data from "Crowdsourcing of Parallel Corpora: the Case of Style Transfer for Detoxification" paper☆14Apr 3, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Multimodal and multilingual topic model with pretrained embeddings☆12Apr 11, 2023Updated 3 years ago
- YOLACT execution with TensorRT☆18Jul 22, 2021Updated 4 years ago
- Fine-tuned Multilingual BERT and Multilingual USE for sentiment analysis in Russian. RuReviews, RuSentiment, Kaggle Russian News Dataset,…☆52Feb 16, 2021Updated 5 years ago
- Compact high quality word embeddings for Russian language☆218Apr 13, 2026Updated 3 months ago
- Демонстрация структуры ml проекта☆11Oct 12, 2022Updated 3 years ago
- this repository is created to accumulate all LaTeX templates needed at Skoltech☆20Nov 27, 2018Updated 7 years ago
- REST API for sentence tokenization and embedding using Multilingual Universal Sentence Encoder.☆51Sep 5, 2021Updated 4 years ago