☆19Aug 3, 2024Updated 2 years ago
Alternatives and similar repositories for FreeEval
Users that are interested in FreeEval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL'24] A Knowledge-grounded Interactive Evaluation Framework for Large Language Models☆40Jul 19, 2024Updated 2 years ago
- Exploiting Unlabeled Data for Target-Oriented Opinion Words Extraction☆24Sep 30, 2022Updated 3 years ago
- Improving fast adversarial training with prior-guided knowledge (TPAMI2024)☆43Apr 21, 2024Updated 2 years ago
- Code for Semantic-Aligned Adversarial Evolution Triangle for High-Transferability Vision-Language Attack(TPAMI 2025)☆42Aug 28, 2025Updated last year
- A Pytorch (support batch and channel) implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech…☆11Jul 24, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A flexible framework for running experiments with PyTorch models in a simulated Federated Learning (FL) environment.☆15Aug 11, 2023Updated 3 years ago
- Improved techniques for optimization-based jailbreaking on large language models (ICLR2025)☆146Apr 7, 2025Updated last year
- Official repository of "Distort, Distract, Decode: Instruction-Tuned Model Can Refine its Response from Noisy Instructions", ICLR 2024 Sp…☆21Mar 7, 2024Updated 2 years ago
- [NDSS'24] Inaudible Adversarial Perturbation: Manipulating the Recognition of User Speech in Real Time☆57Sep 28, 2024Updated last year
- Efficient Dictionary Learning with Switch Sparse Autoencoders (SAEs)☆25Dec 1, 2024Updated last year
- A Japanese G2P tool based on pyopenjtalk☆25Aug 6, 2022Updated 4 years ago
- ☆49Sep 5, 2024Updated last year
- Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning☆39Nov 17, 2024Updated last year
- The repository for paper <Evaluating Open-QA Evaluation>☆25Apr 9, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS'20] Semi-Supervised Partial Label Learning via Confidence-Rated Margin Maximization☆21May 29, 2022Updated 4 years ago
- [ACL 2024]Controlled Text Generation for Large Language Model with Dynamic Attribute Graphs☆40Sep 24, 2024Updated last year
- Code for 🌍 UI-Simulator: LLMs as Scalable, General-Purpose Simulators For Evolving Digital Agent Training☆21Oct 17, 2025Updated 10 months ago
- semi-autoregressive neural machine translation☆23Sep 9, 2018Updated 7 years ago
- SG-Bench: Evaluating LLM Safety Generalization Across Diverse Tasks and Prompt Types☆26Nov 29, 2024Updated last year
- YATO: Yet Another deep learning based Text analysis Open toolkit☆47Oct 11, 2023Updated 2 years ago
- Simple scheduler for running jobs on GPUs☆187Jul 7, 2021Updated 5 years ago
- ☆925May 22, 2024Updated 2 years ago
- you can use dbnet to detect word or bar code,Knowledge Distillation is provided,also python tensorrt inference is provided.☆47Dec 24, 2020Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Build PyTorch CIFAR100 using coarse labels☆39Jun 20, 2020Updated 6 years ago
- The code of paper "Learning to Break the Loop: Analyzing and Mitigating Repetitions for Neural Text Generation" published at NeurIPS 202…☆50Oct 9, 2022Updated 3 years ago
- Official Implementation of "ToolSafe: Enhancing Tool Invocation Safety of LLM-based Agents via Proactive Step-level Guardrail and Feedbac…☆76Mar 25, 2026Updated 5 months ago
- ☆12Apr 21, 2025Updated last year
- javascript animation capture examples 🎬☆13Mar 14, 2023Updated 3 years ago
- LLM benchmarks☆13Feb 22, 2024Updated 2 years ago
- 中文金融大模型测评基准,六大类二十五任务、等级化评价,国内模型获得A级☆10May 6, 2024Updated 2 years ago
- Knowledge Graph based Question Answering benchmark.☆10Feb 1, 2020Updated 6 years ago
- The official GitHub page for the survey paper "A Survey on Evaluation of Large Language Models".☆1,609Aug 1, 2026Updated 3 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Markdown Editor with React + TS + shadcn UI / Tailwind css☆11Jun 3, 2025Updated last year
- Code and data for automatic paraphrase dataset augmentation.☆11Mar 8, 2021Updated 5 years ago
- ☆11Jan 3, 2024Updated 2 years ago
- ☆11Nov 5, 2024Updated last year
- LGEB: Benchmark of Language Generation Evaluation☆16Oct 21, 2022Updated 3 years ago
- Website for release of TellMeWhy dataset for why question answering☆14Nov 11, 2022Updated 3 years ago
- ☆136Oct 16, 2021Updated 4 years ago