Randomized Positional Encodings Boost Length Generalization of Transformers
☆83Mar 14, 2024Updated 2 years ago
Alternatives and similar repositories for randomized_positional_encodings
Users that are interested in randomized_positional_encodings are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Aug 23, 2024Updated 2 years ago
- Exploring Binary Classification Loss for Speaker Verification☆19Jul 18, 2023Updated 3 years ago
- HyPe: Better Pre-trained Language Model Fine-tuning with Hidden Representation Perturbation [ACL 2023]☆14Jul 11, 2023Updated 3 years ago
- Materials of public talks given By SJTU X-LANCE members☆14Dec 3, 2022Updated 3 years ago
- ☆12Jun 5, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Serving Example of CodeGen-350M-Mono-GPTJ on Triton Inference Server with Docker and Kubernetes☆20May 30, 2023Updated 3 years ago
- ☆16Feb 6, 2024Updated 2 years ago
- ☆23Sep 11, 2026Updated 2 weeks ago
- Transformer-based Label Set Generation for Multi-modal Multi-label Emotion Detection☆14Dec 16, 2021Updated 4 years ago
- PPSpeech: Phrase based Parallel End-to-End TTS System☆35Aug 31, 2020Updated 6 years ago
- ☆16Jun 13, 2022Updated 4 years ago
- Understanding Rare Spurious Correlations in Neural Network☆12Jun 5, 2022Updated 4 years ago
- This repo contains the code associated to the paper: "Constrained Causal Bayesian Optimization" by Aglietti Virginia, Alan Malek, Ira Kt …☆16Jun 18, 2024Updated 2 years ago
- ☆12Sep 1, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- mPLM-Sim: Better Cross-Lingual Similarity and Transfer in Multilingual Pretrained Language Models☆11Jan 19, 2024Updated 2 years ago
- Synthetic data generator for machine learning☆16Oct 18, 2023Updated 2 years ago
- ☆26Apr 11, 2023Updated 3 years ago
- ☆22Oct 3, 2024Updated last year
- Lightning Attention-2: A Free Lunch for Handling Unlimited Sequence Lengths in Large Language Models☆344Feb 23, 2025Updated last year
- ☆20Apr 2, 2025Updated last year
- Repo for "Zemi: Learning Zero-Shot Semi-Parametric Language Models from Multiple Tasks" ACL 2023 Findings☆15May 3, 2023Updated 3 years ago
- Revisiting Efficient Training Algorithms For Transformer-based Language Models (NeurIPS 2023)☆81Aug 30, 2023Updated 3 years ago
- Curriculum training☆24Jun 25, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Spherical residual vector quantization (SRVQ)☆31Aug 25, 2024Updated 2 years ago
- [ICLR 2024] CLEX: Continuous Length Extrapolation for Large Language Models☆78Mar 12, 2024Updated 2 years ago
- ☆266Feb 25, 2020Updated 6 years ago
- PyTorch source code of NAACL 2021 paper "Improving the Lexical Ability of Pretrained Language Models for Unsupervised Neural Machine Tran…☆18Oct 18, 2022Updated 3 years ago
- ☆28Apr 24, 2026Updated 5 months ago
- Extending context length of visual language models☆12Dec 18, 2024Updated last year
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated 2 years ago
- Ring attention implementation with flash attention☆1,060Sep 10, 2025Updated last year
- Speech Separation☆21Mar 7, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Running inference on the ZeroSCROLLS benchmark☆22Apr 18, 2024Updated 2 years ago
- Sequence modeling with Mega.☆301Jan 28, 2023Updated 3 years ago
- A research project and comparative study on various Active Noise Cancellation Algorithms like FxLMS, EMFN, Chebyshev filter and Hammerste…☆10Jul 3, 2022Updated 4 years ago
- Efficient Personalized Speech Enhancement through Self-Supervised Learning☆24Mar 12, 2023Updated 3 years ago
- annotated-transformer-kr☆15May 16, 2019Updated 7 years ago
- [AAAI 2024] Code for CTX-vec2wav in UniCATS☆130Jun 11, 2024Updated 2 years ago
- Easy-to-use framework for evaluating cross-lingual consistency of factual knowledge (Supported LLaMA, BLOOM, mT5, RoBERTa, etc.) Paper he…☆28Aug 8, 2025Updated last year