☆39Mar 25, 2026Updated 4 months ago
Alternatives and similar repositories for FANformer
Users that are interested in FANformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository hosts the source code for the paper "ROCODE: Integrating Backtracking Mechanism and Program Analysis in Large Language Mo…☆16Dec 16, 2025Updated 7 months ago
- ☆16Nov 26, 2024Updated last year
- ☆264Oct 26, 2025Updated 9 months ago
- Xmixers: A collection of SOTA efficient token/channel mixers☆29Sep 4, 2025Updated 10 months ago
- ☆27Aug 31, 2025Updated 10 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆13Apr 15, 2024Updated 2 years ago
- ☆14Nov 20, 2022Updated 3 years ago
- Official implementation of the paper "Pretraining Language Models to Ponder in Continuous Space"☆26Jul 21, 2025Updated last year
- ☆10Dec 17, 2024Updated last year
- Spectral Attention Autoregressive Model (SAAM)☆17Oct 27, 2022Updated 3 years ago
- Agent Skill for ROS/ROS2 robot control via rosbridge WebSocket.☆24Feb 27, 2026Updated 5 months ago
- Efficient PScan implementation in PyTorch☆17Jan 2, 2024Updated 2 years ago
- Code for the paper: https://arxiv.org/pdf/2309.06979.pdf☆21Jul 29, 2024Updated 2 years ago
- PyTorch implementation for HyperMixing, a linear-time token-mixing technique used in HyperMixer architecture☆26Jun 12, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year
- Flash-Linear-Attention models beyond language☆21Aug 28, 2025Updated 11 months ago
- Tiled Flash Linear Attention library for fast and efficient mLSTM Kernels.☆91Jul 6, 2026Updated 3 weeks ago
- 使用django+pyecharts+PP-Human开发的动态数据大屏, 有人流数据的采集入库, 打架、摔倒等事件警报,口罩检测等实用功能。边缘端版本使用onnx推理提升效率,服务端版本支持视频流推拉☆34May 3, 2023Updated 3 years ago
- Official repo for "ProSec: Fortifying Code LLMs with Proactive Security Alignment"☆18Feb 26, 2026Updated 5 months ago
- Official repository of paper "RNNs Are Not Transformers (Yet): The Key Bottleneck on In-context Retrieval"☆27Apr 17, 2024Updated 2 years ago
- 这是本科毕业设计的课题,“基于深度网络的网站验证码识别研究与实现”。主要是利用卷积神经网络,基于TensorFlow平台,构建了三层卷积两层全联接模型,训练出的一个准确率为91.3%的识别模型。再基于Django构建登陆系统,使用selenium实现自动测试,完成验证码从识…☆25Jun 18, 2018Updated 8 years ago
- [COLING 2020] BERT-based Models for Chengyu☆17Dec 29, 2021Updated 4 years ago
- ☆94Oct 11, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The test of different distributed-training methods on High-Flyer AIHPC☆27Oct 18, 2022Updated 3 years ago
- ☆14Jun 24, 2024Updated 2 years ago
- Source code for CoNLL 2021 paper by Huebner et al. 2021☆21Jul 13, 2023Updated 3 years ago
- Official Code Repository for the paper "Key-value memory in the brain"☆32Feb 25, 2025Updated last year
- Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence☆68Nov 11, 2025Updated 8 months ago
- Code for ICML 2024 paper☆34Sep 18, 2025Updated 10 months ago
- [ICML 24 NGSM workshop] Associative Recurrent Memory Transformer implementation and scripts for training and evaluation☆67Mar 12, 2026Updated 4 months ago
- ☆35Apr 12, 2024Updated 2 years ago
- DOMAINEVAL is an auto-constructed benchmark for multi-domain code generation that consists of 2k+ subjects (i.e., description, reference …☆13Dec 12, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- JoinAI是一个开源仓库,专注于算法工程能力的培养,包括工程和数学原理的整理☆11Apr 20, 2025Updated last year
- Collection of red team scripts, resources & configs.☆15Feb 14, 2026Updated 5 months ago
- ☆13Mar 21, 2023Updated 3 years ago
- MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆27May 23, 2026Updated 2 months ago
- The official repository for our paper "The Neural Data Router: Adaptive Control Flow in Transformers Improves Systematic Generalization".☆34Jun 11, 2025Updated last year
- ☆27Jun 7, 2026Updated last month
- A pytorch implementation of a text to videos GAN☆12Jul 26, 2019Updated 7 years ago