We view Large Language Models as stochastic language layers in a network, where the learnable parameters are the natural language prompts at each layer. We stack two such layers, feeding the output of one layer to the next. We call the stacked architecture a Deep Language Network - DLN
☆95Jul 25, 2024Updated 2 years ago
Alternatives and similar repositories for deep-language-networks
Users that are interested in deep-language-networks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR'24 Spotlight] DP-OPT: Make Large Language Model Your Privacy-Preserving Prompt Engineer☆48May 30, 2024Updated 2 years ago
- Byte-sized text games for code generation tasks on virtual environments☆20Jul 8, 2024Updated 2 years ago
- ☆26Apr 17, 2024Updated 2 years ago
- EMNLP'2022: BERTScore is Unfair: On Social Bias in Language Model-Based Metrics for Text Generation☆41Oct 19, 2022Updated 3 years ago
- Text Adventure Learning Environment Suite - Benchmark to evaluate language models on interactive text environments.☆32Sep 10, 2026Updated last month
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [NeurIPS 2025] Official repo of "Martian World Model: Controllable Video Synthesis with Physically Accurate 3D Reconstructions"☆21Aug 6, 2025Updated last year
- Building modular LMs with parameter-efficient fine-tuning.☆116Sep 15, 2026Updated 3 weeks ago
- ☆13Dec 13, 2022Updated 3 years ago
- Official code for FAccT'21 paper "Fairness Through Robustness: Investigating Robustness Disparity in Deep Learning" https://arxiv.org/abs…☆13Mar 9, 2021Updated 5 years ago
- Code for pre-training BabyLM baseline models.☆16Jun 19, 2023Updated 3 years ago
- Template-DQN and DRRN agent implementations☆23Jun 12, 2023Updated 3 years ago
- ☆18Oct 12, 2022Updated 3 years ago
- A Data Source for Reasoning Embodied Agents☆20Sep 18, 2023Updated 3 years ago
- NumGLUE: A Suite of Fundamental yet Challenging Mathematical Reasoning Tasks☆20May 10, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICML 2024 Spotlight] Differentially Private Synthetic Data via Foundation Model APIs 2: Text☆61Jan 11, 2025Updated last year
- Project for SNARE benchmark☆11Jun 5, 2024Updated 2 years ago
- ☆30Jun 19, 2023Updated 3 years ago
- source code for ICLR'24 paper "How does unlabeled data provably help OOD detection?"☆14Feb 1, 2024Updated 2 years ago
- Accompanying repo for the RLPrompt paper☆367Jun 6, 2024Updated 2 years ago
- A framework for human-readable prompt-based method with large language models. Specially designed for researchers. (Deprecated, check out…☆130Feb 25, 2023Updated 3 years ago
- This repository contains some of the code used in the paper "Training Language Models with Langauge Feedback at Scale"☆27Mar 30, 2023Updated 3 years ago
- This is a new metric that can be used to evaluate faithfulness of text generated by LLMs. The work behind this repository can be found he…☆31Aug 25, 2023Updated 3 years ago
- ☆46Apr 10, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Counterfactual Evaluation and Learning for Interactive Systems: Foundations, Implementations, and Recent Advances☆12Aug 14, 2022Updated 4 years ago
- Code Release for "Broken Neural Scaling Laws" (BNSL) paper☆59Oct 29, 2023Updated 2 years ago
- awesome-LLM-controlled-constrained-generation☆57Aug 16, 2024Updated 2 years ago
- A full-text error corrector for English based on transformers and deep learning☆10Jan 8, 2023Updated 3 years ago
- context denoising training for long-context modeling☆17Oct 10, 2025Updated last year
- Data and info for the paper "ParaDetox: Text Detoxification with Parallel Data"☆34Apr 2, 2025Updated last year
- [ICML 2023] "Robust Weight Signatures: Gaining Robustness as Easy as Patching Weights?" by Ruisi Cai, Zhenyu Zhang, Zhangyang Wang☆16May 4, 2023Updated 3 years ago
- ☆10Dec 12, 2023Updated 2 years ago
- Code for experiments for "ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy"☆102Sep 11, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Super fast implementations of common benchmark text world games☆55Aug 25, 2025Updated last year
- Code for RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs. ACL 2023.☆63Nov 27, 2024Updated last year
- Mixed integer programming for computing lipschitz constants of ReLU Networks☆17Feb 10, 2023Updated 3 years ago
- Fine-tuning, DPO, RLHF, RLAIF on LLMs - Qwen3, Zephyr 7B GPTQ with 4-Bit Quantization, Mistral-7B-GPTQ☆15Jul 5, 2025Updated last year
- Ask Me Anything language model prompting☆548Jul 5, 2023Updated 3 years ago
- CliniDeID automatically de-identifies clinical text notes according to the HIPAA Safe Harbor method. It accurately finds identifiers and …☆11Aug 13, 2023Updated 3 years ago
- In this repository, I place my solution for the exercises in multiple famous math textbooks, including Stochastic Differential Equation, …☆14Nov 13, 2023Updated 2 years ago