Optimus is a flexible and scalable framework built to train language models efficiently across diverse hardware configurations, including CPU, AMD, and NVIDIA GPUs.
β70Dec 4, 2025Updated 7 months ago
Alternatives and similar repositories for EuroBERT
Users that are interested in EuroBERT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NLP with Rust for Python π¦πβ72Jun 9, 2026Updated last month
- One-stop shop for running and fine-tuning transformer-based language models for retrievalβ65Jul 9, 2026Updated last week
- Datamodels for hugging face tokenizersβ108Jun 18, 2026Updated last month
- My configuration files, loosely inspired by @sontekβ39Jul 6, 2026Updated 2 weeks ago
- Train Models Contrastively in Pytorchβ798Mar 26, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The Solr Package Directory and Sanctuaryβ13May 28, 2026Updated last month
- Chunk Dedupe Estimationβ20Nov 5, 2024Updated last year
- Prompt Jinja2 templates for LLMsβ36Jul 9, 2025Updated last year
- β11Oct 11, 2023Updated 2 years ago
- Baguetter is a flexible, efficient, and hackable search engine library implemented in Python. It's designed for quickly benchmarking, impβ¦β210Aug 31, 2024Updated last year
- GLiNER model in a FastAPI microservice.β47Dec 11, 2024Updated last year
- β77May 14, 2026Updated 2 months ago
- The NLP Bias Identification Toolkitβ39Sep 8, 2023Updated 2 years ago
- Code, results and other artifacts from the paper introducing the WildChat-50m dataset and the Re-Wild model family.β38Apr 1, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Bringing BERT into modernity via both architecture changes and scalingβ1,701Mar 1, 2026Updated 4 months ago
- Source codes and datasets for paper "Leveraging Transferable Knowledge Concept Graph Embedding for Cold-Start Cognitive Diagnosis" (SIGIRβ¦β21Mar 15, 2024Updated 2 years ago
- β53Feb 10, 2025Updated last year
- Benchmarking LLM Inference Speedsβ14May 17, 2026Updated 2 months ago
- A minimal hackable implementation of policy gradient methods (GRPO, PPO, REINFORCE)β16Feb 20, 2026Updated 5 months ago
- Lightweight continuous batching OpenAI compatibility using HuggingFace Transformers include T5 and Whisper.β29Mar 15, 2025Updated last year
- Code for KaLM-Embedding modelsβ116Jun 30, 2025Updated last year
- β63Jan 26, 2025Updated last year
- High-Performance Engine for Multi-Vector Searchβ268May 28, 2026Updated last month
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- CodeSage: Code Representation Learning At Scale (ICLR 2024)β122Oct 27, 2024Updated last year
- Build text classifiers using 3 most popular machine learning / deep learning frameworks - Scikit-learn, PyTorch, TensorFlowβ10Sep 22, 2021Updated 4 years ago
- Visual demo of DSPy's prompt optimization on Gradioβ15Apr 14, 2025Updated last year
- Tailscale fixesβ12Sep 17, 2024Updated last year
- Template for getting started with Hybrid Dagster Cloudβ15Sep 19, 2025Updated 10 months ago
- implementation of https://arxiv.org/pdf/2312.09299β21Jul 3, 2024Updated 2 years ago
- Exploitability calculation for imperfect-information game benchmarksβ37Apr 5, 2025Updated last year
- Recipes for learning, fine-tuning, and adapting ColPali to your multimodal RAG use cases. π¨π»βπ³β356Jun 2, 2025Updated last year
- Computational Modelling of Social Systems at the University of Konstanzβ11Jun 3, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Extract full next-token probabilities via language model APIsβ248Feb 23, 2024Updated 2 years ago
- code for training & evaluating Contextual Document Embedding modelsβ207May 14, 2025Updated last year
- Open-source RL Framework with Online Teacher-Student Distillationβ22Mar 5, 2026Updated 4 months ago
- Dataset of conversations, generated by prompting Gemini Ultra. These are conversations between a teacher and a student, where the teacherβ¦β35Oct 29, 2024Updated last year
- Private Prompts Prototype Documentationβ38Oct 29, 2025Updated 8 months ago
- β108Jun 2, 2025Updated last year
- Repository for the code of the "PPL-MCTS: Constrained Textual Generation Through Discriminator-Guided Decoding" paper, NAACL'22β66Oct 25, 2022Updated 3 years ago