☆37Oct 1, 2025Updated 11 months ago
Alternatives and similar repositories for FoNE
Users that are interested in FoNE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- On Efficient Language and Vision Assistants for Visually-Situated Natural Language Understanding: What Matters in Reading and Reasoning, …☆20Mar 13, 2026Updated 6 months ago
- Tensorflow Custom Callbacks in Custom Training Loop☆10Apr 8, 2021Updated 5 years ago
- Causal Inference for Time Series Data (with CausalML Demo)☆14Jun 11, 2023Updated 3 years ago
- Clustered Compositional Embeddings☆13Oct 25, 2023Updated 2 years ago
- Ἀνατομή is a PyTorch library to analyze representation of neural networks☆13Jan 31, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- codes and plots for "Active-Dormant Attention Heads: Mechanistically Demystifying Extreme-Token Phenomena in LLMs"☆11Dec 30, 2024Updated last year
- ☆14Oct 12, 2024Updated last year
- ☆12Sep 16, 2024Updated 2 years ago
- Official PyTorch implementation of Extract Free Dense Misalignment from CLIP (AAAI'25)☆24Apr 20, 2025Updated last year
- ☆12Aug 25, 2026Updated 3 weeks ago
- ☆11Jun 19, 2024Updated 2 years ago
- GIANT (Gene-based data Integration and ANalysis Technique) is a method for large-scale joint analyses of atlas-level single cell data.☆14Jun 13, 2023Updated 3 years ago
- [NeurIPS2022] Where to Pay Attention in Sparse Training for Feature Selection?☆13Feb 10, 2023Updated 3 years ago
- Find context neurons in Pythia models.☆13Jun 13, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆29Oct 6, 2024Updated last year
- [ECCV24] "Challenging Forgets: Unveiling the Worst-Case Forget Sets in Machine Unlearning" by Chongyu Fan*, Jiancheng Liu*, Alfred Hero, …☆28May 27, 2025Updated last year
- [NeurIPS'25] Towards Interpretability Without Sacrifice: Faithful Dense Layer Decomposition with Mixture of Decoders☆16May 28, 2025Updated last year
- Least Squares Regression for subspace clustering☆11May 27, 2018Updated 8 years ago
- The official implementation of HybridNorm: Towards Stable and Efficient Transformer Training via Hybrid Normalization☆19Mar 7, 2025Updated last year
- [COLM 2024] Large Language Models as Biomedical Hypothesis Generators: A Comprehensive Evaluation☆15Jul 15, 2024Updated 2 years ago
- Implementation of Unified Embedding: Battle-Tested Feature Representations for Web-Scale ML Systems☆15Nov 11, 2023Updated 2 years ago
- Codes for the paper The emergence of clusters in self-attention dynamics.☆17Dec 18, 2023Updated 2 years ago
- A tool for visualising kmers in 2D space.☆14Feb 23, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The Hessian of tall-skinny networks is easy to invert☆17Sep 5, 2026Updated 2 weeks ago
- a light structured Residual autoencoder and mutual nearest neighbor Paring guided Adversarial Network for scRNA-seq batch correction☆15Jul 13, 2023Updated 3 years ago
- (CVPR 2024) FLHetBench: Benchmarking Device and State Heterogeneity in Federated Learning☆19Jun 21, 2024Updated 2 years ago
- Dual Adversarial Autoencoder for Generating Set-valued Sequences☆19Jan 15, 2021Updated 5 years ago
- The full training script for Enformer - Tensorflow Sonnet☆18Apr 18, 2022Updated 4 years ago
- A clustering algorithm that can perform internal validation inspired by forest fire dynamics and self-organized criticality☆18May 18, 2022Updated 4 years ago
- UIUC CS 225 SP20 ZJUI Environment☆12May 30, 2020Updated 6 years ago
- [TrimKV] Cache What Lasts: Token Retention for Memory-Bounded KV Cache in LLMs - [DBTrimKV] Make Each Token Count: Towards Improving Lo…☆21Jul 26, 2026Updated last month
- Large Language Models Can Be Contextual Privacy Protection Learners☆18Oct 28, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official implementation of "Data Mixture Inference: What do BPE tokenizers reveal about their training data?"☆23May 15, 2025Updated last year
- [ICML25] Official repo for "Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond…☆25Sep 27, 2025Updated 11 months ago
- ☆58Nov 24, 2024Updated last year
- Pretrain, finetune and deploy AI models on multiple GPUs, TPUs with zero code changes.☆13Feb 20, 2024Updated 2 years ago
- ☆12Sep 7, 2026Updated 2 weeks ago
- ☆21Jul 25, 2023Updated 3 years ago
- Code for PAC-Bayes Compression Bounds So Tight That They Can Explain Generalization, NeurIPS 2022☆18Nov 23, 2022Updated 3 years ago