☆25Sep 24, 2024Updated last year
Alternatives and similar repositories for w2s
Users that are interested in w2s are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LLM play 20questions with itself☆13Mar 31, 2023Updated 3 years ago
- The code of paper "Learning to Break the Loop: Analyzing and Mitigating Repetitions for Neural Text Generation" published at NeurIPS 202…☆50Oct 9, 2022Updated 3 years ago
- [ICML 2024] "Envisioning Outlier Exposure by Large Language Models for Out-of-Distribution Detection"☆15Feb 15, 2025Updated last year
- ☆22Feb 10, 2025Updated last year
- ☆15May 14, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- An archive of learning resources assembled by current Exun members and alumni.☆15Jun 23, 2026Updated 2 months ago
- Uncertainty-Aware Curriculum Learning for Neural Machine Translation (ACL 2020)☆11Jun 12, 2020Updated 6 years ago
- NeuroSurgeon is a package that enables researchers to uncover and manipulate subnetworks within models in Huggingface Transformers☆44Feb 12, 2025Updated last year
- ☆11Aug 10, 2024Updated 2 years ago
- Yet another dynamic batch sampler for variable sequence data in PyTorch.☆13Dec 9, 2021Updated 4 years ago
- LLM RL envs done right, plus some training code☆17Feb 1, 2026Updated 7 months ago
- [ICLR 2025] Code&Data for the paper "Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization"☆15Jun 21, 2024Updated 2 years ago
- ☆12Mar 4, 2025Updated last year
- Official project page for Estimating the Rate-Distortion Function by Wasserstein Gradient Descent☆19Nov 2, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- EMNLP 2020: Filtering before Iteratively Referring for Knowledge-Grounded Response Selection in Retrieval-Based Chatbots☆12Dec 15, 2020Updated 5 years ago
- Official implementation of ICML 2025 paper "Understanding Multimodal LLMs Under Distribution Shifts: An Information-Theoretic Approach"☆12May 27, 2025Updated last year
- Code for CVPR paper: Computationally Budgeted Continual Learning: What Does Matter?☆17Mar 16, 2024Updated 2 years ago
- Code for preprint: Summarizing Differences between Text Distributions with Natural Language☆43Feb 24, 2023Updated 3 years ago
- Co-Supervised Learning: Improving Weak-to-Strong Generalization with Hierarchical Mixture of Experts☆15Feb 26, 2024Updated 2 years ago
- ☆20May 23, 2025Updated last year
- ☆22Jan 5, 2024Updated 2 years ago
- The code of “Improving Weak-to-Strong Generalization with Scalable Oversight and Ensemble Learning”☆17Feb 26, 2024Updated 2 years ago
- Code and Dataset release of "Carpe Diem: On the Evaluation of World Knowledge in Lifelong Language Models" (NAACL 2024)☆10Oct 16, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆24Dec 17, 2025Updated 8 months ago
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 6 months ago
- A Public repository for the COMeT model☆14Jul 25, 2024Updated 2 years ago
- ☆15Jul 14, 2022Updated 4 years ago
- Code and data for the paper "Emergent Visual-Semantic Hierarchies in Image-Text Representations" (ECCV 2024)☆34Jul 27, 2026Updated last month
- Released Apertium translation pairs☆32May 27, 2021Updated 5 years ago
- ☆15Nov 3, 2022Updated 3 years ago
- 🔥This is a repository of paper list for streaming LLMs/MLLMs.☆28Apr 19, 2026Updated 4 months ago
- Comparison of gradient estimation techniques for black-box adversarial examples☆11Oct 31, 2018Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Implementation of Boundary Attributions for Normal (Vector) Explanations☆11Aug 13, 2021Updated 5 years ago
- Cookbooks showcasing various applications of Cleanlab☆22Jan 20, 2026Updated 7 months ago
- Source code for ScaleGrad☆19Dec 28, 2021Updated 4 years ago
- Proof-of-concept of global switching between numpy/jax/pytorch in a library.☆17Jun 18, 2024Updated 2 years ago
- Identification of the Adversary from a Single Adversarial Example (ICML 2023)☆10Jul 15, 2024Updated 2 years ago
- James' cookbook of evaluations and finetuning experiments☆35Feb 19, 2026Updated 6 months ago
- We are the 21-day expandables of a kaggle competition.☆15Jul 23, 2017Updated 9 years ago