[ICML'25] "Rethinking Addressing in Language Models via Contextualized Equivariant Positional Encoding" by Jiajun Zhu, Peihao Wang, Ruisi Cai, Jason D. Lee, Pan Li, Zhangyang Wang
☆15Jun 6, 2025Updated last year
Alternatives and similar repositories for TAPE
Users that are interested in TAPE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025] 🔍 Multilingual Evaluation of English-Centric LLMs via Cross-Lingual Alignment☆11Apr 6, 2025Updated last year
- PathPiece tokenizer☆14Nov 10, 2024Updated last year
- [ICML 2023] "Data Efficient Neural Scaling Law via Model Reusing" by Peihao Wang, Rameswar Panda, Zhangyang Wang☆14Jan 4, 2024Updated 2 years ago
- Curriculum training☆24Jun 25, 2025Updated last year
- DP-GNN design that ensures both model weights and inference procedure differentially private (NeurIPS 2023)☆12Oct 14, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [VLDB'23] SUREL+ is a novel set-based computation framework for scalable subgraph-based graph representation learning.☆17Apr 10, 2025Updated last year
- [ICML 2024] Code for Pairwise Alignment Improves Graph Domain Adaptation (Pair-Align)☆14Jun 15, 2024Updated 2 years ago
- [Preprint] Graph State Space Convolution (GSSC)☆14Jun 11, 2024Updated 2 years ago
- The official repository for Toxic Commons and Celadon. Toxicity Classification for public domain data.☆23Jul 17, 2026Updated 2 months ago
- The repository for 'Unsupervised Learning for Combinatorial Optimization with Principled Proxy Design'☆16Oct 9, 2022Updated 3 years ago
- [NeurIPS 2024] | An Efficient Recipe for Long Context Extension via Middle-Focused Positional Encoding☆23Oct 10, 2024Updated last year
- Code for GeSS: Benchmarking Geometric Deep Learning under Scientific Applications with Distribution Shifts☆16Dec 28, 2024Updated last year
- Multilingual Knowledge Graph Enhancement (EMNLP 2023)☆25Sep 11, 2026Updated last week
- Small python package to measure OCR quality and other related metrics.☆26Feb 19, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Subset-Norm and Subset-Momentum. This repo is built on top of https://github.com/jiaweizzhao/GaLore.☆19Jul 9, 2025Updated last year
- Code and data for the paper "Turning English-centric LLMs Into Polyglots: How Much Multilinguality Is Needed?"☆26Jun 3, 2025Updated last year
- Reinforcement Learning with Pong in the Browser via TensorFlow.js☆17Jan 4, 2023Updated 3 years ago
- ☆18Oct 12, 2025Updated 11 months ago
- [NeurIPS 2022] "Signal Processing for Implicit Neural Representations" by Dejia Xu*, Peihao Wang*, Yifan Jiang, Zhiwen Fan, Zhangyang Wan…☆53Dec 1, 2022Updated 3 years ago
- ☆20Oct 25, 2022Updated 3 years ago
- German Language Understanding Evaluation Benchmark @NAACL24☆23Dec 11, 2025Updated 9 months ago
- Official implementation of "Can Test-Time Scaling Improve World Foundation Model?"☆15Jul 12, 2025Updated last year
- [VLDB'22] SUREL is a novel walk-based computation framework for efficient subgraph-based graph representation learning.☆20Apr 10, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICML 2023] Structural Re-weighting Improves Graph Domain Adaptation (StruRW)☆23Jun 20, 2023Updated 3 years ago
- ☆18Jun 30, 2025Updated last year
- [ICML 2024 Oral] LSH-Based Efficient Point Transformer (HEPT)☆27Jan 24, 2025Updated last year
- [ICML 2025] Official code of "AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization"☆32Jan 10, 2026Updated 8 months ago
- xLM is a modular, research-friendly framework for developing and comparing non-autoregressive language models. Built on PyTorch and PyTor…☆31Updated this week
- ☆13Dec 2, 2024Updated last year
- 基于中心度的中文关键短语抽取工具☆11Sep 2, 2022Updated 4 years ago
- [NeurIPS'24 LanGame workshop] On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability☆42Apr 10, 2026Updated 5 months ago
- The open source implementation of the multi grouped query attention by the paper "GQA: Training Generalized Multi-Query Transformer Model…☆17Dec 11, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆12Dec 13, 2022Updated 3 years ago
- ☆15Nov 20, 2025Updated 10 months ago
- Linear Attention for Efficient Bidirectional Sequence Modeling☆18May 13, 2025Updated last year
- Code for the CVPR '23 paper, "Defending Against Patch-based Backdoor Attacks on Self-Supervised Learning"☆10Jun 9, 2023Updated 3 years ago
- ☆18Sep 22, 2024Updated last year
- KG data for ODA☆12May 14, 2026Updated 4 months ago
- [ICML 2025] Generalization Principles for Inference over Text-Attributed Graphs with Large Language☆22Jul 15, 2025Updated last year