Official implementation of "BERTs are Generative In-Context Learners"
☆32Mar 14, 2025Updated last year
Alternatives and similar repositories for bert-in-context
Users that are interested in bert-in-context are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An AI character interaction system with emotional modeling and advanced memory management☆17Aug 6, 2026Updated 3 weeks ago
- A compact high-signal benchmark for evaluating frontier agents☆27Aug 3, 2026Updated 3 weeks ago
- Trully flash implementation of DeBERTa disentangled attention mechanism.☆91Feb 10, 2026Updated 6 months ago
- ☆23Jun 18, 2026Updated 2 months ago
- ☆12Jan 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Cuda implemenation of flash-kmeans, 2x faster☆23May 8, 2026Updated 3 months ago
- ☆16May 14, 2024Updated 2 years ago
- Deep learning models to predict enhancers in different Drosophila embryo tissues☆20Dec 10, 2023Updated 2 years ago
- Chainer and PyTorch implementation of GAN with gradient reversal layer☆10Mar 19, 2022Updated 4 years ago
- Evolution-inspired data augmentations for PyTorch-based models for regulatory genomics☆25Jun 3, 2025Updated last year
- ☆11Feb 9, 2024Updated 2 years ago
- It is about how to load and aggregate pretrained word embeddings in pytorch, e.g., ELMo\BERT\XLNET.☆12Mar 2, 2020Updated 6 years ago
- Chu-Lui-Edmonds decoding extracted from TurboParser☆14May 16, 2017Updated 9 years ago
- ☆22Oct 14, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code repository for study ''Evaluating the representational power of pre-trained DNA language models for regulatory genomics"☆26Jun 26, 2024Updated 2 years ago
- KernelBench v2: Can LLMs Write GPU Kernels? - Benchmark with Torch -> Triton (and more!) problems☆25Jul 4, 2025Updated last year
- Train a SmolLM-style llm on fineweb-edu in JAX/Flax with an assortment of optimizers.☆19Jul 24, 2025Updated last year
- Morfessor EM+Prune☆10Jul 22, 2020Updated 6 years ago
- ☆13Jan 9, 2022Updated 4 years ago
- ☆15Jun 26, 2026Updated 2 months ago
- ☆10Mar 4, 2025Updated last year
- Notebooks for training universal 0-shot classifiers on many different tasks☆141Dec 28, 2024Updated last year
- LinPipe: Multilingual Processing Tool☆10Jan 20, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ACL21 Math Word Problem Solving with Explicit Numerical Values☆13Nov 10, 2021Updated 4 years ago
- Simple Scalable Discrete Diffusion for text in PyTorch☆37Sep 27, 2024Updated last year
- ☆11Feb 28, 2025Updated last year
- GraphPart, a data partitioning method for ML on biological sequences☆36Oct 26, 2023Updated 2 years ago
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notation☆14Jan 2, 2026Updated 8 months ago
- Your favourite classical machine learning algos on the GPU/TPU☆23Dec 14, 2025Updated 8 months ago
- Implementation of Variational Auto-Encoder for text generation in pytorch.☆12Oct 9, 2020Updated 5 years ago
- Code for EMNLP 2021 Paper "Recall and Learn: A Memory-augmented Solver for Math Word Problems".☆16Oct 20, 2022Updated 3 years ago
- Fine-tune ModernBERT with custom tokenizers, curriculum learning, and next-gen optimizers.☆74Jan 16, 2026Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Equivariant layers for RC-complement symmetry in DNA sequence data☆13Feb 24, 2022Updated 4 years ago
- ☆14Jul 22, 2021Updated 5 years ago
- Code for Teacher-Student Networks with Multiple Decoders for Solving Math Word Problem (IJCAI 2020).☆11Sep 19, 2020Updated 5 years ago
- ☆12Nov 16, 2023Updated 2 years ago
- Code for ACL 2023 Paper: ACLM: A Selective-Denoising based Generative Data Augmentation Approach for Low-Resource Complex NER☆22Jul 19, 2023Updated 3 years ago
- Efficient encoder-decoder architecture for small language models (≤1B parameters) with cross-architecture knowledge distillation and visi…☆32Feb 7, 2025Updated last year
- A deep learning-based predictor of enzyme optimal pH☆12Jun 25, 2025Updated last year