Preserving Generalization of Language Models in Few-shot Continual Relation Extraction (EMNLP2024)
☆17Nov 21, 2024Updated last year
Alternatives and similar repositories for FCRE-via-MMI
Users that are interested in FCRE-via-MMI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- FHBI: Improving Generalization with Flat Hilbert Bayesian Inference (ICML2025)☆19Jan 7, 2026Updated 8 months ago
- Official implementation of "From Implicit to Explicit Feedback: A deep neural network for modeling sequential behaviours and long-short t…☆19Oct 16, 2025Updated 11 months ago
- ☆22Sep 3, 2024Updated 2 years ago
- Pre-training script for BART in JAX/Flax☆37Aug 4, 2022Updated 4 years ago
- Planning for Success: Exploring LLM Long-term Planning Capabilities in Table Understanding☆17Jun 17, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Dataset and Code for NeurIPS 2023 paper "Language-driven Scene Synthesis using Multi-conditional Diffusion Model."☆47Aug 8, 2024Updated 2 years ago
- DocChecker: Bootstrapping Code-Text Pretrained Language Model to Detect Inconsistency Between Code and Comment☆15Jan 23, 2024Updated 2 years ago
- A novel variant of sliced Wasserstein based on a new slicing technique that utilizes the convolution operator.☆12Jan 14, 2023Updated 3 years ago
- Official codes for EMNLP 2024 paper "Multi-expert Prompting Improves Reliability, Safety and Usefulness of Large Language Models"☆39Dec 14, 2024Updated last year
- ☆14Feb 22, 2022Updated 4 years ago
- [EMNLP 2023] The Vault: A Comprehensive Multilingual Dataset for Advancing Code Understanding and Generation☆105Aug 21, 2024Updated 2 years ago
- ☆14Dec 22, 2024Updated last year
- Stochastic Multiple Target Sampling Gradient Descent (NeurIPS 2022)☆13Sep 19, 2022Updated 4 years ago
- [TrimKV] Cache What Lasts: Token Retention for Memory-Bounded KV Cache in LLMs - [DBTrimKV] Make Each Token Count: Towards Improving Lo…☆21Jul 26, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implementation of "Mixture of Experts Meets Prompt-Based Continual Learning" (NeurIPS 2024)☆43Aug 1, 2025Updated last year
- ☆16Jan 28, 2024Updated 2 years ago
- Code for NeurIPS 2021 paper "Flattening Sharpness for Dynamic Gradient Projection Memory Benefits Continual Learning".☆16Oct 18, 2021Updated 4 years ago
- This is an implementation of Graph Convolution Topic Model for Data Streams☆15Sep 12, 2022Updated 4 years ago
- ☆16May 10, 2026Updated 4 months ago
- Official PyTorch implementation for the paper Minimizing Trajectory Curvature of ODE-based Generative Models, ICML 2023☆91Feb 14, 2025Updated last year
- ☆29Jun 2, 2026Updated 3 months ago
- Joint Distillation for Fast Likelihood Evaluation and Sampling in Flow-based Models☆23Mar 5, 2026Updated 6 months ago
- Code to replicate the Representation Noising paper and tools for evaluating defences against harmful fine-tuning☆24Dec 12, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆22Aug 19, 2025Updated last year
- [ICLR 2025] CAMEx: Curvature-Aware Merging of Experts☆24Mar 1, 2025Updated last year
- ☆18Apr 28, 2023Updated 3 years ago
- 👀 Visual Instruction Inversion: Image Editing via Visual Prompting (NeurIPS 2023)☆96Dec 19, 2023Updated 2 years ago
- [NeurIPS'24] Protecting Your LLMs with Information Bottleneck☆25Nov 7, 2024Updated last year
- The official implementation of T3D: T3D: Few-Step Diffusion Language Models via Trajectory Self-Distillation with Direct Discriminative O…☆26Jul 10, 2026Updated 2 months ago
- official code repo for paper "Merging Models on the Fly Without Retraining: A Sequential Approach to Scalable Continual Model Merging"☆25Oct 11, 2025Updated 11 months ago
- D-JEPA on ImageNet☆29Nov 18, 2024Updated last year
- Simple Guidance Mechanisms for Discrete Diffusion Models☆93Dec 16, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆37Feb 13, 2024Updated 2 years ago
- PyTorch implementation of Drifting Models by Kaiming He et al.☆22Feb 6, 2026Updated 7 months ago
- [NeurIPS 24] Alleviating Distortion in Image Generation via Multi-Resolution Diffusion Models☆45Sep 30, 2024Updated last year
- An efficient distillation method for flow matching models☆31Feb 1, 2026Updated 7 months ago
- [CVPR2024] Official Codes for "Adversarial Score Distillation: When score distillation meets GAN"☆38Apr 24, 2025Updated last year
- ☆46Jun 15, 2023Updated 3 years ago
- [CVPR 2024] When StyleGAN Meets Stable Diffusion: a W+ Adapter for Personalized Image Generation☆161Aug 18, 2025Updated last year