The code and data for the GPT-4 based benchmark in the vicuna blog post
☆42Aug 2, 2023Updated 2 years ago
Alternatives and similar repositories for vicuna-blog-eval
Users that are interested in vicuna-blog-eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Code for NAACL 2022 paper: "Persona-Guided Planning for Controlling the Protagonist's Persona in Story Generation"☆16Sep 1, 2022Updated 3 years ago
- [EMNLP 2024] RoLoRA: Fine-tuning Rotated Outlier-free LLMs for Effective Weight-Activation Quantization☆41Sep 24, 2024Updated last year
- ☆13Feb 2, 2021Updated 5 years ago
- Code for the paper "Rethinking Benchmark and Contamination for Language Models with Rephrased Samples"☆325Dec 20, 2023Updated 2 years ago
- Unofficial Scalable-Softmax Is Superior for Attention☆21May 30, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementation of "Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model Architectures"☆15Jan 21, 2026Updated 6 months ago
- ☆26Dec 2, 2024Updated last year
- torch_quantizer is a out-of-box quantization tool for PyTorch models on CUDA backend, specially optimized for Diffusion Models.☆25Mar 29, 2024Updated 2 years ago
- mini-swe-agent-plus: a tiny (~100 LOC) GitHub issue fixer—now with a robust multi-line text edit tool.☆25Jan 20, 2026Updated 6 months ago
- 🏆 The winner code for Neurips'23 BigANN Competition OOD and Sparse track.☆15Jun 17, 2025Updated last year
- UI for ActivityWatch. Include category editor and viewer for multiple categorizations.☆10Jan 31, 2024Updated 2 years ago
- Benchmarking Attention Mechanism in Vision Transformers.☆20Oct 10, 2022Updated 3 years ago
- ☆10Jun 28, 2023Updated 3 years ago
- Text generation using language models with multiple exit heads☆16Sep 18, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Help creating image dataset for machine learning.☆10Nov 4, 2020Updated 5 years ago
- ☆17May 10, 2024Updated 2 years ago
- ☆10Dec 18, 2023Updated 2 years ago
- Training with Block Minifloat number representation☆18May 2, 2021Updated 5 years ago
- Chrome Extension. As the name suggests.☆10Jan 30, 2022Updated 4 years ago
- (TMLR J2C Certification) Fed-SB: A Silver Bullet for Extreme Communication Efficiency and Performance in (Private) Federated LoRA Fine-Tu…☆27Oct 4, 2025Updated 9 months ago
- ☆14May 21, 2024Updated 2 years ago
- Python package for rematerialization-aware gradient checkpointing☆27Oct 31, 2023Updated 2 years ago
- Pandas Helper Library for reading and writing DataFrames from and to HBase.☆10Mar 8, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- codes and plots for "Active-Dormant Attention Heads: Mechanistically Demystifying Extreme-Token Phenomena in LLMs"☆11Dec 30, 2024Updated last year
- ☆18Jan 20, 2026Updated 6 months ago
- [ACM MM 25] FingER: Content Aware Fine-grained Evaluation with Reasoning for AI-Generated Videos☆17Jul 17, 2025Updated last year
- An Attention Superoptimizer☆22Jan 20, 2025Updated last year
- A3C tensorflow implementation☆11Jul 22, 2018Updated 8 years ago
- ☆23Nov 7, 2025Updated 8 months ago
- Japanese semantic test suite (FraCaS counterpart and extensions)☆13Apr 21, 2026Updated 3 months ago
- ☆14Jul 13, 2025Updated last year
- Official PyTorch implementation of Elastic-InfoGAN [NeurIPS 2020]☆15Apr 3, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- "Can images help recognize entities? A study of the role of images for Multimodal NER" (W-NUT at EMNLP 2021)☆21Nov 14, 2021Updated 4 years ago
- ☆11Mar 28, 2021Updated 5 years ago
- ☆234Jun 11, 2024Updated 2 years ago
- Code of fine-tuning neural sparse models and training from scratch. #SIGIR2025☆26Mar 11, 2026Updated 4 months ago
- The official implementation of the EMNLP 2023 paper LLM-FP4☆226Dec 15, 2023Updated 2 years ago
- BESA is a differentiable weight pruning technique for large language models.☆17Mar 4, 2024Updated 2 years ago
- {DeepL, Google, WMT-Best, davinci-003, turbo, gpt-4} × {En-De, En-Cs, En-Ru, En-Zh, De-Fr, En-Ja, Uk-En, Uk-Cs, En-Hr, En-Ha, En-Is}☆14Jun 18, 2023Updated 3 years ago