☆51Sep 28, 2025Updated last year
Alternatives and similar repositories for elephant
Users that are interested in elephant are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Findings of EMNLP 2025] Benchmark for evaluating sycophantic behavior in multi-turn, free-form conversational settings.☆38Dec 19, 2025Updated 9 months ago
- datasets from the paper "Towards Understanding Sycophancy in Language Models"☆137Oct 25, 2023Updated 2 years ago
- code and data associated with CoMPosT: Characterizing and Evaluating Caricature in LLM Simulations☆11Oct 13, 2023Updated 2 years ago
- Augmenting Statistical Models with Natural Language Parameters☆28Sep 17, 2024Updated 2 years ago
- ⚔️ OpenHands PR Arena ⚔️ is a platform for evaluating and benchmarking agentic coding assistants through paired pull request (PR) generat…☆18Dec 15, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Providing the answer to "How to do patching on all available SAEs on GPT-2?". It is an official repository of the implementation of the p…☆14Jan 26, 2025Updated last year
- Metric Space Magnitude Computations☆15Jun 30, 2026Updated 2 months ago
- Code associated with the paper: Neural Representations for Modeling Variation in Speech.☆18Mar 10, 2022Updated 4 years ago
- Streamlit AI story generator with multi-step LLM prompting, local OpenAI-compatible models, and OpenRouter support.☆17Jun 18, 2026Updated 3 months ago
- ☆24Mar 8, 2024Updated 2 years ago
- Code for the creation of the SPoRC dataset☆21Nov 18, 2024Updated last year
- ☆14Jun 7, 2023Updated 3 years ago
- A free, research-based vocab web app for beginners of Chinese with Django backend.☆10Feb 19, 2024Updated 2 years ago
- Code for evaluating AI systems on the MASK honesty benchmark.☆27Mar 6, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆39Jun 25, 2026Updated 3 months ago
- Code of paper: Probing the Difficulty Perception Mechanism of Large Language Models☆19Mar 17, 2026Updated 6 months ago
- playing with gpt4☆13Mar 17, 2023Updated 3 years ago
- ☆53Dec 2, 2025Updated 9 months ago
- ☆15May 15, 2026Updated 4 months ago
- Benchmark to estimate model sycophancy☆34Nov 30, 2025Updated 9 months ago
- The offical code for paper "What Constitutes a Faithful Summary? Preserving Author Perspectives in News Summarization"☆10Jun 23, 2024Updated 2 years ago
- URS Benchmark: Evaluating LLMs on User Reported Scenarios☆31May 30, 2025Updated last year
- Data, codebook, and models to automatically detect storytelling.☆30Apr 23, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆14Aug 30, 2023Updated 3 years ago
- CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation☆14Aug 19, 2025Updated last year
- Code and data for Marked Personas (ACL 2023)☆30May 26, 2023Updated 3 years ago
- ☆20May 1, 2025Updated last year
- [NeurIPS XAIA & Springer] Code and notebooks to paper "A Fresh Look at Sanity Checks for Saliency Maps"☆25Jul 12, 2024Updated 2 years ago
- The code implementation of the paper Stumbling Blocks: Stress Testing the Robustness of Machine-Generated Text Detectors Under Attacks (A…☆13Jul 16, 2024Updated 2 years ago
- Custom importer for node-sass to import packages from the `node_modules` directory.☆11Aug 8, 2017Updated 9 years ago
- Official source code for the paper "Tailored Design of Audio-Visual Speech Recognition Models using Branchformers"☆15Feb 24, 2025Updated last year
- Code for "FactKB: Generalizable Factuality Evaluation using Language Models Enhanced with Factual Knowledge". EMNLP 2023.☆20Dec 25, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Exploring aspects of similarity between spoken personal narratives by disentangling them into narrative clause types -- Supplementary inf…☆12Jul 14, 2020Updated 6 years ago
- Repository for the WACV 2024 paper "PsyMo: A Dataset for Estimating Self-Reported Psychological Traits from Gait"☆14Feb 22, 2024Updated 2 years ago
- ICLR 2025 Workshop & CHI 2025 SIG: "Bidirectional Human-AI Alignment"☆59Aug 6, 2024Updated 2 years ago
- ☆35Nov 7, 2024Updated last year
- ☆30Aug 2, 2024Updated 2 years ago
- This repository includes the code implementation of the paper Improving Pacing in Long-Form Story Planning by Yichen Wang, Kevin Yang, Xi…☆18Nov 19, 2024Updated last year
- Resolving Knowledge Conflicts in Large Language Models, COLM 2024☆18Oct 7, 2025Updated 11 months ago