Official implementation of "BERTs are Generative In-Context Learners"
☆32Mar 14, 2025Updated last year
Alternatives and similar repositories for bert-in-context
Users that are interested in bert-in-context are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An AI character interaction system with emotional modeling and advanced memory management☆17Aug 6, 2026Updated last week
- Official implementation of "GPT or BERT: why not both?"☆64Jul 28, 2025Updated last year
- Code for the paper "Function-Space Learning Rates"☆23Jun 3, 2025Updated last year
- Natural language understanding benchmarks for Norwegian☆14Aug 29, 2025Updated 11 months ago
- Trully flash implementation of DeBERTa disentangled attention mechanism.☆91Feb 10, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆22Jun 18, 2026Updated last month
- ☆12Jan 2, 2024Updated 2 years ago
- The repo of the Doc2SoarGraph framework☆10Sep 17, 2024Updated last year
- ☆10Sep 11, 2020Updated 5 years ago
- ☆16May 14, 2024Updated 2 years ago
- Deep learning models to predict enhancers in different Drosophila embryo tissues☆20Dec 10, 2023Updated 2 years ago
- Evolution-inspired data augmentations for PyTorch-based models for regulatory genomics☆25Jun 3, 2025Updated last year
- It is about how to load and aggregate pretrained word embeddings in pytorch, e.g., ELMo\BERT\XLNET.☆12Mar 2, 2020Updated 6 years ago
- Chu-Lui-Edmonds decoding extracted from TurboParser☆14May 16, 2017Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- LTG-Bert☆35Jan 8, 2024Updated 2 years ago
- ☆22Oct 14, 2024Updated last year
- Code repository for study ''Evaluating the representational power of pre-trained DNA language models for regulatory genomics"☆26Jun 26, 2024Updated 2 years ago
- KernelBench v2: Can LLMs Write GPU Kernels? - Benchmark with Torch -> Triton (and more!) problems☆24Jul 4, 2025Updated last year
- Train a SmolLM-style llm on fineweb-edu in JAX/Flax with an assortment of optimizers.☆19Jul 24, 2025Updated last year
- MISCA: A Joint Model for Multiple Intent Detection and Slot Filling with Intent-Slot Co-Attention (EMNLP 2023 - Findings)☆35Jul 22, 2024Updated 2 years ago
- Scaling Laws for Mixture of Experts Models☆15Feb 25, 2025Updated last year
- ☆13Jan 9, 2022Updated 4 years ago
- ☆34Jan 25, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- MultiLexNorm 2021 competition system from ÚFAL☆16Dec 30, 2021Updated 4 years ago
- LinPipe: Multilingual Processing Tool☆10Jan 20, 2026Updated 6 months ago
- Code base of In-Context Learning for Dialogue State tracking☆45Sep 24, 2023Updated 2 years ago
- ACL21 Math Word Problem Solving with Explicit Numerical Values☆13Nov 10, 2021Updated 4 years ago
- ☆11Feb 28, 2025Updated last year
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notation☆14Jan 2, 2026Updated 7 months ago
- Your favourite classical machine learning algos on the GPU/TPU☆23Dec 14, 2025Updated 8 months ago
- LLM training in simple, raw C/CUDA☆15Dec 5, 2024Updated last year
- Code for EMNLP 2021 Paper "Recall and Learn: A Memory-augmented Solver for Math Word Problems".☆16Oct 20, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆10Jul 10, 2023Updated 3 years ago
- Equivariant layers for RC-complement symmetry in DNA sequence data☆13Feb 24, 2022Updated 4 years ago
- ☆14Jul 22, 2021Updated 5 years ago
- Code for Teacher-Student Networks with Multiple Decoders for Solving Math Word Problem (IJCAI 2020).☆11Sep 19, 2020Updated 5 years ago
- Efficient encoder-decoder architecture for small language models (≤1B parameters) with cross-architecture knowledge distillation and visi…☆32Feb 7, 2025Updated last year
- Collection of LLM completions for reasoning-gym task datasets☆31Jul 4, 2025Updated last year
- ☆14Aug 28, 2022Updated 3 years ago