☆15Dec 29, 2025Updated 7 months ago
Alternatives and similar repositories for agi-eval
Users that are interested in agi-eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A iterative feedback driven benchmark on LLM's instruction following ability☆58May 25, 2026Updated 2 months ago
- ☆11Nov 16, 2019Updated 6 years ago
- WBench: A Comprehensive Multi-turn Benchmark for Interactive Video World Model Evaluation☆181Jul 31, 2026Updated 2 weeks ago
- Lyric: A Rust-powered secure runtime for AI-Agent.☆24Mar 11, 2025Updated last year
- deepFM 実装☆14Aug 9, 2017Updated 9 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR'26] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆26May 9, 2026Updated 3 months ago
- IPA files available for download from ONEJailbreak.com☆13Apr 30, 2025Updated last year
- workflows for alfred4☆14Nov 16, 2020Updated 5 years ago
- reviese pyrouge files for supporting winxp win 8.1 win10☆12Nov 21, 2017Updated 8 years ago
- ☆14Oct 30, 2021Updated 4 years ago
- Bayesian Deep Active Learning for Named entity recognition (NER)☆19Jan 17, 2020Updated 6 years ago
- Commonsense Knowledge Base Reasoning☆10Sep 3, 2018Updated 7 years ago
- ☆115Jul 3, 2026Updated last month
- ☆42Apr 7, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Gives cross platform support for multiline selectable text, where no keyboard is wanted.☆18Dec 15, 2023Updated 2 years ago
- ☆20May 24, 2024Updated 2 years ago
- [Neurips 2022] “ Back Razor: Memory-Efficient Transfer Learning by Self-Sparsified Backpropogation”, Ziyu Jiang*, Xuxi Chen*, Xueqin Huan…☆19Mar 14, 2023Updated 3 years ago
- A comparison of pretraining framework for LLM☆22Feb 6, 2025Updated last year
- [NAACL 2022] "Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training", Yuanxin Liu, Fandong Meng, Zheng Lin, Pe…☆15Oct 18, 2022Updated 3 years ago
- ☆470Jul 21, 2026Updated 3 weeks ago
- ☆20Mar 30, 2022Updated 4 years ago
- ☆20Dec 16, 2020Updated 5 years ago
- Tensorflow-LSTM-CRF tool for Named Entity Recognizer☆59Jul 23, 2017Updated 9 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [EMNLP 2023]Context Compression for Auto-regressive Transformers with Sentinel Tokens☆25Nov 6, 2023Updated 2 years ago
- This is the libMF source files with comments in Chinses.☆30May 25, 2014Updated 12 years ago
- GoJieba Bleve support☆27May 8, 2026Updated 3 months ago
- Code for "A Unified Model for Joint Chinese Word Segmentation and Dependency Parsing"☆39May 24, 2022Updated 4 years ago
- BiLSTM-CRF for sequence labeling in Dynet☆82Jun 15, 2017Updated 9 years ago
- Implementation of the cw2vec model☆29Jul 20, 2018Updated 8 years ago
- DDRel: A new dataset for interpersonal relation classification in dyadic dialogues☆23Sep 12, 2021Updated 4 years ago
- ☆24Jun 13, 2023Updated 3 years ago
- [EMNLP'24] LongHeads: Multi-Head Attention is Secretly a Long Context Processor☆32Apr 8, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Multi-turn response selection using dialogue dependency relations☆24Sep 1, 2021Updated 4 years ago
- A flagship 560-billion-parameter open-source MoE model that advances Native Formal Reasoning in Lean4 for Mathematics Formalization and P…☆94May 9, 2026Updated 3 months ago
- [AAAI 2021]Knowledge-Driven Distractor Generation for Cloze-Style Multiple Choice Questions☆22Jul 29, 2021Updated 5 years ago
- Code for the paper "Knowledge-driven Data Construction for Zero-shot Evaluation in Commonsense Question Answering" (AAAI 2021)☆30Feb 19, 2021Updated 5 years ago
- Implementation of "LM-Infinite: Simple On-the-Fly Length Generalization for Large Language Models"☆40Nov 11, 2024Updated last year
- ☆50Jun 7, 2025Updated last year
- Extremely Long-Horizon Agentic Tasks Requiring Active Acting and Inductive Reasoning☆34Feb 9, 2026Updated 6 months ago