☆58Jul 8, 2025Updated last year
Alternatives and similar repositories for CSInternship2025
Users that are interested in CSInternship2025 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2024] Wrong-of-Thought: An Integrated Reasoning Framework with Multi-Perspective Verification and Wrong Information☆13Oct 1, 2024Updated last year
- ☆34Jul 8, 2025Updated last year
- ☆68Aug 7, 2026Updated last month
- code for "CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models"☆19Mar 10, 2025Updated last year
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generation☆48Jul 22, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- 🔥Awesome Multimodal Large Language Models Paper List☆154Mar 12, 2025Updated last year
- The Source Code for OmniVideoBench @ICLR 2026☆78Feb 12, 2026Updated 7 months ago
- Nanjing University Advanced Machine Learning Review☆32Jun 11, 2025Updated last year
- Official Implementation of "UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and Generation"☆143Oct 17, 2025Updated 11 months ago
- Official implementation of StableI2I (ICML 2026)☆20Aug 28, 2026Updated 3 weeks ago
- VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆18Feb 3, 2026Updated 7 months ago
- Interleaving Reasoning: Next-Generation Reasoning Systems for AGI☆283Jun 5, 2026Updated 3 months ago
- [Survey] Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey☆478Jan 17, 2025Updated last year
- A curated list of models, benchmarks, tools and guides for audio editing☆46Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official GitHub repo for Scaling Physical Reasoning with the PHYSICS Dataset (NeurIPS25).☆16Sep 20, 2025Updated last year
- ☆23Apr 11, 2026Updated 5 months ago
- iFSQ & LlamaGen-REPA☆106Jan 27, 2026Updated 7 months ago
- The Source Code for IF-VidCap @ICLR 2026☆18Oct 22, 2025Updated 11 months ago
- MVU-Eval @NeurIPS DB 2025☆16Nov 11, 2025Updated 10 months ago
- An official implementation of "SPARK: Synergistic Policy And Reward Co-Evolving Framework"☆25Oct 23, 2025Updated 11 months ago
- Demo-ICL: In-Context Learning for Procedural Video Knowledge Acquisition☆47Mar 3, 2026Updated 6 months ago
- A roadmap of artificial intelligence☆17Sep 17, 2022Updated 4 years ago
- [INTERSPEECH 2026 Oral]Official code for "Semantic-VAE: Semantic-Alignment Latent Representation for Better Speech Synthesis"☆125Jun 21, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A tool for better use of Inspire platform (Beta: Codeberg version is more up-to-date)☆30Apr 2, 2026Updated 5 months ago
- 使用自然语言绘制流程图,基于OpenAI☆12Nov 13, 2023Updated 2 years ago
- (ICLR 2025 Spotlight) Official code repository for Interleaved Scene Graph.☆31Aug 7, 2025Updated last year
- ☆40Jul 21, 2026Updated 2 months ago
- 2024年保研经验贴和相关物料☆1,078Jul 7, 2024Updated 2 years ago
- Official Implementation of "What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion"☆83May 27, 2026Updated 3 months ago
- [ISSTA 2025] A Large-scale Empirical Study on Fine-tuning Large Language Models for Unit Testing☆13Feb 9, 2025Updated last year
- [INTERSPEECH 2025 Oral]Official code for "Accelerating Diffusion-based Text-to-Speech Model Training with Dual Modality Alignment"☆68Jun 16, 2025Updated last year
- ☆111Sep 11, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ✨✨ [ICLR 2026] MME-Unify: A Comprehensive Benchmark for Unified Multimodal Understanding and Generation Models☆43Apr 10, 2025Updated last year
- 🦞 ClawMark: A Living-World Benchmark for Multi-Day, Multimodal Coworker Agents☆123May 28, 2026Updated 3 months ago
- 操作系统课程设计:FAT文件系统☆12Mar 14, 2025Updated last year
- Official Implementation of "Visual-ERM: Reward Modeling for Visual Equivalence"☆66Mar 23, 2026Updated 6 months ago
- Explore how to get a VQ-VAE models efficiently!☆69Jul 24, 2025Updated last year
- TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs☆136Jul 27, 2026Updated last month
- An AI-powered content conversion tool that transforms text, web content, or HTML code into beautifully designed card images.一款基于AI的内容转换工…☆35Jul 29, 2025Updated last year