[ACL 2023]: Training Trajectories of Language Models Across Scales https://arxiv.org/pdf/2212.09803.pdf
☆25Nov 14, 2023Updated 2 years ago
Alternatives and similar repositories for training_trajectory_analysis
Users that are interested in training_trajectory_analysis are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CHIL 2024] Interpretation of Intracardiac Electrograms Through Textual Representations☆12Sep 4, 2024Updated 2 years ago
- Evaluating Durability: Benchmark Insights into Multimodal Watermarking☆12Jun 7, 2024Updated 2 years ago
- Code and dataset for Polyglot Prompting: Multilingual Multitask Prompt Training.☆18Dec 7, 2022Updated 3 years ago
- ☆14Oct 28, 2023Updated 2 years ago
- Language Model Baselines for PyTorch☆41Aug 18, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ECCV 2022] "Improve Few-Shot Transfer Learning with Low-Rank Decompose and Align" by Ziyu Jiang, Tianlong Chen, Xuxi Chen, Yu Cheng, Luo…☆13Jul 19, 2022Updated 4 years ago
- ☆10Jul 16, 2023Updated 3 years ago
- Structured Pruning Adapters in PyTorch☆19Aug 30, 2023Updated 3 years ago
- Functional Optimal Transport: Map Estimation and Domain Adaptation for Functional data☆28Jun 7, 2021Updated 5 years ago
- LoCoBench-Agent: An Interactive Benchmark for LLM Agents in Long-Context Software Engineering☆22Jun 2, 2026Updated 3 months ago
- [EACL 2023] Transfer Knowledge from Natural Language to Electrocardiography: Can We Detect Cardiovascular Disease Through Language Models…☆18May 7, 2024Updated 2 years ago
- Official code for the paper Improving Language Plasticity via Pretraining with Active Forgetting, NeurIPS 2023☆21Mar 12, 2026Updated 5 months ago
- Code accompanying our paper at AISTATS 2020☆21Jan 12, 2021Updated 5 years ago
- Maze navigation with MLM-U☆17Dec 21, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆10Aug 18, 2016Updated 10 years ago
- Source code and data for The Magic of IF: Investigating Causal Reasoning Abilities in Large Language Models of Code (Findings of ACL 2023…☆31Jun 4, 2023Updated 3 years ago
- Code for Deep learning models for electrocardiograms are susceptible to adversarial attack☆23Feb 4, 2021Updated 5 years ago
- DiWA: Diverse Weight Averaging for Out-of-Distribution Generalization☆31Jan 31, 2023Updated 3 years ago
- [EMNLP 2023] An Empirical Exploration of Cross-domain Alignment between Language and Electroencephalogram☆31Nov 9, 2023Updated 2 years ago
- ☆25Mar 4, 2022Updated 4 years ago
- Pre-trained models for our work on Temporal Graph Generation☆19Jun 2, 2021Updated 5 years ago
- Fun project to run your own LLM chat bot using llama.cpp☆11Jun 9, 2023Updated 3 years ago
- REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy Transfer (ICML 2022 Long Oral)☆27Sep 10, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ACL 2022] Structured Pruning Learns Compact and Accurate Models https://arxiv.org/abs/2204.00408☆198May 9, 2023Updated 3 years ago
- We conduct a preregistered experiment to investigate whether fact checks provided by a large language model can serve as an effective mis…☆13Dec 14, 2024Updated last year
- Semantic Parser Localizer (SPL) code repository☆10Mar 15, 2021Updated 5 years ago
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 3 years ago
- BLOOM+1: Adapting BLOOM model to support a new unseen language☆75Mar 2, 2024Updated 2 years ago
- TyDiP Multilingual Politeness dataset and code☆12Oct 15, 2023Updated 2 years ago
- [DMLR 2024] Benchmarking Robustness of Multimodal Image-Text Models under Distribution Shift☆40Jan 25, 2024Updated 2 years ago
- Preprint: Asymmetry in Low-Rank Adapters of Foundation Models☆40Feb 27, 2024Updated 2 years ago
- ☆13Oct 8, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This is a comprehensive guide on how you can automate your feature engineering process.☆11Jun 25, 2018Updated 8 years ago
- ☆12Sep 1, 2023Updated 3 years ago
- ☆10Oct 28, 2019Updated 6 years ago
- Evaluating Visual Fidelity of Image Descriptions☆11Aug 15, 2019Updated 7 years ago
- 🧬 BioRelEx: Biological Relation Extraction Benchmark @ ACL BioNLP Workshop 2019☆18Jul 31, 2019Updated 7 years ago
- Official code repository for the main conference paper in ACL2023: COLA: Contextualized Commonsense Causality Reasoning from the Causal I…☆34May 12, 2023Updated 3 years ago
- Code for the NIPS 2016 paper "Single-Image Depth Perception in the Wild"☆11Nov 17, 2017Updated 8 years ago