This is the official repo for the paper "General365: Benchmarking General Reasoning in LLMs under High Difficulty and Diversity".
☆88Apr 14, 2026Updated 3 months ago
Alternatives and similar repositories for General365
Users that are interested in General365 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official repo for the paper "AMO-Bench: Large Language Models Still Struggle in High School Math Competitions".☆177Feb 6, 2026Updated 5 months ago
- [ACL2026] UCAS: Uncertainty-aware Advantage Shaping for RLVR☆31Apr 14, 2026Updated 3 months ago
- MemoryDial☆15Mar 10, 2026Updated 4 months ago
- [ACL26 Findings] TopoDIM: One-shot Topology Generation of Diverse Interaction Modes for Multi-Agent Systems☆19Jan 19, 2026Updated 6 months ago
- This is the official repository for JailExpert☆23Sep 9, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Data and Code Repository for “STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems”☆17Apr 17, 2026Updated 3 months ago
- [ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning☆18Apr 17, 2026Updated 3 months ago
- [NeurIPS 2025] SYMPHONY: Synergistic Multi-agent Planning with Heterogeneous Language Model Assembly☆17Oct 22, 2025Updated 9 months ago
- Prototype Conditioned Generative Replay for Continual Learning in NLP - NAACL 2025☆26Updated this week
- [🏆CVPR'26] Official Repo for IAG: Input-aware Backdoor Attack on VLM-based Visual Grounding☆33Jun 2, 2026Updated 2 months ago
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 3 months ago
- Official Codebase of the ACL 2026 Oral paper "Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contra…☆26Jun 25, 2026Updated last month
- Implementation for paper Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs, which is accepted by ACL 2026 (main con…☆16Oct 10, 2025Updated 9 months ago
- videoPro: Adaptive Program Reasoning for Long Video Understanding☆45Apr 15, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆17Feb 22, 2026Updated 5 months ago
- ☆79Apr 12, 2026Updated 3 months ago
- [Paper][EMNLP 2025] Enrich-on-Graph: Query-Graph Alignment for Complex Reasoning with LLM Enriching☆35Feb 8, 2026Updated 5 months ago
- [ACL 2026] Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue☆24Apr 14, 2026Updated 3 months ago
- [🏆AAAI'25] Official Repo for ChemVLM: Exploring the Power of Multimodal Large Language Models in Chemistry Area.☆88Apr 14, 2026Updated 3 months ago
- [AAAI2024] Debiasing Multimodal Sarcasm Detection with Contrastive Learning☆17Jan 5, 2024Updated 2 years ago
- This repository contains the official implementation of the paper "BioGraphFusion: Graph Knowledge Embedding for Biological Completion an…☆16Nov 5, 2025Updated 8 months ago
- [ICLR'26] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆27May 9, 2026Updated 2 months ago
- ☆19Jun 26, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Benchmarking Language Agents Under Controllable and Extreme Context Growth☆51Apr 29, 2026Updated 3 months ago
- A flagship 560-billion-parameter open-source MoE model that advances Native Formal Reasoning in Lean4 for Mathematics Formalization and P…☆93May 9, 2026Updated 2 months ago
- ☆52Jun 12, 2026Updated last month
- The implementation of ACL main 2026 paper "ReCreate: Reasoning and Creating Domain Agents Driven by Experience"☆164Apr 29, 2026Updated 3 months ago
- IntentVCNet: Bridging Spatio-Temporal Gaps for Intention-Oriented Controllable Video Captioning☆19Aug 16, 2025Updated 11 months ago
- Decoupled Gradient Policy Optimization (DGPO) - Official Implementation☆48Apr 22, 2026Updated 3 months ago
- Revealing the Unstable Foundations of eBPF-Based Kernel Extensions☆18May 20, 2025Updated last year
- A Network Integration Approach for Drug-Target Interaction Prediction☆13Apr 5, 2025Updated last year
- ☆14Oct 7, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆30Apr 29, 2026Updated 3 months ago
- ☆14Jul 28, 2025Updated last year
- ☆13Apr 30, 2025Updated last year
- ☆13May 31, 2023Updated 3 years ago
- [CVPR 2025] LION-FS: Fast & Slow Video-Language Thinker as Online Video Assistant☆29Dec 2, 2025Updated 8 months ago
- On Path to Multimodal Generalist: General-Level and General-Bench☆21Jul 11, 2025Updated last year
- Synthetic pretraining data by rephrasing the web☆26Jun 5, 2026Updated last month