NeurIPS 2025 Poster
☆26Feb 4, 2025Updated last year
Alternatives and similar repositories for Adaptive_Distractions
Users that are interested in Adaptive_Distractions are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NeurIPS 2025 Poster☆23Oct 17, 2025Updated 10 months ago
- ProbeLLM: Automating Principled Diagnosis of LLM Failures☆17Feb 11, 2026Updated 6 months ago
- [NeurIPS 2024] HonestLLM: Toward an Honest and Helpful Large Language Model☆29Jun 10, 2025Updated last year
- [ICLR'25] DataGen: Unified Synthetic Dataset Generation via Large Language Models☆69Mar 8, 2025Updated last year
- [ICLR'26, NAACL'25 Demo] Toolkit & Benchmark for evaluating the trustworthiness of generative foundation models.☆135Aug 22, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PostgreSQL extension which allows to translate a given source SQL statement into another pre-defined SQL statement.☆24Sep 25, 2025Updated 10 months ago
- Code for the 2025 ACL publication "Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs"☆32Jun 25, 2025Updated last year
- [ICLR'26] Building a Foundational Guardrail for General Agentic Systems via Synthetic Data☆49Oct 26, 2025Updated 9 months ago
- [ACM MM'24] CRLD: Cross-View Consistency Regularisation for Knowledge Distillation☆15Apr 10, 2025Updated last year
- TextPy: Collaborative Agent Workflow through Programming and Prompting☆27May 9, 2025Updated last year
- [ICLR'24] MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use☆118Mar 21, 2024Updated 2 years ago
- [ACL 2026] A benchmark for evaluating the reliability of text-to-infographic generation with curated test cases and automated question-ba…☆15Jun 8, 2026Updated 2 months ago
- This repository is the official implementation for VISD.☆23May 17, 2026Updated 3 months ago
- [ICML 2024 Oral] Official code repository for MLLM-as-a-Judge.☆95Feb 17, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official Implementation for: "RAW: A Robust and Agile Plug-and-Play Watermark Framework for AI-Generated Images (Videos) with Provable Gu…☆38Oct 30, 2024Updated last year
- [ICCV 2025] MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation☆23Sep 5, 2025Updated 11 months ago
- Benchmarking data and script used for LLM multi-agent collaboration systems from AWS Bedrock Agents Science team.☆18Dec 10, 2024Updated last year
- Fleming-VL: Towards Universal Medical Visual Understanding with Multimodal LLMs☆15Nov 6, 2025Updated 9 months ago
- Learned Query Optimizer☆13Mar 16, 2022Updated 4 years ago
- 本项目分享了本人在四川大学计算机学院计算机科学与技术专业的各类课程的资料、学习建议以及作业。欢迎使用,也希望其他校友能为此库提供缺失资料,如果喜欢就Star吧。☆10May 18, 2021Updated 5 years ago
- Chat about anything on any video!☆39Sep 5, 2023Updated 2 years ago
- ☆15May 15, 2026Updated 3 months ago
- Operating Systems Internals and Design principles 8th 读书笔记,资源整理☆19Jun 3, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- C++ 实现 BP 神经网络识别手写数字数据集 MNIST☆14Jan 7, 2024Updated 2 years ago
- ☆12May 6, 2022Updated 4 years ago
- TPC-DS Generation, Execution and Analyzer for Postgres☆20Dec 6, 2022Updated 3 years ago
- Can We Trust Large Language Models?: A Benchmark for Responsible Large Language Models via Toxicity, Bias, and Value-alignment Evaluation☆25Oct 12, 2023Updated 2 years ago
- The official repository for "MemSim: A Bayesian Simulator for Evaluating Memory of LLM-based Personal Assistants".☆17Oct 10, 2024Updated last year
- ☆11Dec 23, 2024Updated last year
- MedGo: Medical Large Language Model Based on Qwen3-32B☆21Dec 3, 2025Updated 8 months ago
- Official Repo for SvS: A Self-play with Variational Problem Synthesis strategy for RLVR training☆56Dec 13, 2025Updated 8 months ago
- [ICML 2024] TrustLLM: Trustworthiness in Large Language Models☆630Jun 24, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Fine-tuning, DPO, RLHF, RLAIF on LLMs - Qwen3, Zephyr 7B GPTQ with 4-Bit Quantization, Mistral-7B-GPTQ☆15Jul 5, 2025Updated last year
- [CVPR 2024] Narrative Action Evaluation with Prompt-Guided Multimodal Interaction☆43May 16, 2024Updated 2 years ago
- Language Models as Multi-Modal Query Planners☆21Mar 20, 2024Updated 2 years ago
- ☆22May 21, 2025Updated last year
- ☆10Jun 29, 2020Updated 6 years ago
- Monitor and navigate AI coding agent sessions in tmux with an fzf picker, status widget, and crash detection.☆30Jul 15, 2026Updated last month
- The official repo for "Ref-AVS: Refer and Segment Objects in Audio-Visual Scenes", ECCV 2024☆50Oct 12, 2025Updated 10 months ago