☆35May 12, 2026Updated 4 months ago
Alternatives and similar repositories for Phi-Ground
Users that are interested in Phi-Ground are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Home page for Microsoft Phi-Ground tech-report☆22Sep 8, 2025Updated last year
- Official Repo for MageBench: Bridging Large Multimodal Models to Agents☆21Jan 8, 2025Updated last year
- This repo is for CaesarNeRF: Calibrated Semantic Representation for Few-Shot Generalizable Neural Rendering.☆14Mar 6, 2024Updated 2 years ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆23Apr 24, 2026Updated 5 months ago
- PyTorch implementation of the article "Generative Adversarial Network for Handwritten Text"☆10Nov 13, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆13Apr 2, 2024Updated 2 years ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 5 months ago
- Official repository for Find n' Propagate: Open-Vocabulary 3D Object Detection in Urban Environments☆18Jul 9, 2024Updated 2 years ago
- ☆11Jan 17, 2021Updated 5 years ago
- [AAAI 2025] Official Implementation of I-HallA v1.0☆16Feb 2, 2025Updated last year
- [CVPR 2025] OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts☆20Apr 2, 2025Updated last year
- Quickly hashing all subexpressions of a program modulo alpha-renaming☆18Sep 7, 2021Updated 5 years ago
- To appear in the 11th International Conference on Learning Representations (ICLR 2023).☆18Feb 24, 2023Updated 3 years ago
- Azure AI Visual Search toolkit☆15Oct 25, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ECML 2020] OBProx-SG☆16Mar 23, 2021Updated 5 years ago
- [ACL'25 (Findings)] Explorer: Scaling Exploration-driven Web Trajectory Synthesis for Multimodal Web Agents☆29Feb 17, 2026Updated 7 months ago
- Implementation of Balanced Graph Partitioning Konstantin" - Andreev and Harald Racke (Authors of the paper) by Ivan Vigorito and Lorenzo …☆14Feb 17, 2023Updated 3 years ago
- ☆37Apr 1, 2026Updated 6 months ago
- Workshop content for our first MCP workshop☆13Mar 20, 2025Updated last year
- Learn how to efficiently fine-tune large language models (LLMs) in the OCI Generative AI playground.☆11Nov 13, 2024Updated last year
- A dedicated space for developer experience☆17Jan 5, 2026Updated 9 months ago
- replace the current round robin scheduler in xv6 with a lottery scheduler☆13Oct 19, 2019Updated 6 years ago
- [ACM MM25] LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models☆25Mar 29, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Jan 26, 2023Updated 3 years ago
- Official implementation of Latent-GRPO: reinforcement learning for vocabulary-space latent reasoning.☆21Aug 31, 2026Updated last month
- ☆40Jul 15, 2025Updated last year
- A Retrieval-Augmented Generation (RAG) system running DeepSeek R1 Distill LLama 70B model using Groq's fast inference API.☆13Jan 29, 2025Updated last year
- TaiYiXLCheckpointLoader: An unoffical node support Taiyi-Diffusion-XL(Taiyi-XL) Chinese-English bilingual language model☆10Sep 1, 2024Updated 2 years ago
- Accurate Retraining-free Pruning for Pretrained Encoder-based Language Models (ICLR 2024)☆14May 31, 2025Updated last year
- PresentAgent-2: Towards Generalist Multimodal Presentation Agents☆18Jun 5, 2026Updated 4 months ago
- ☆49Feb 10, 2026Updated 7 months ago
- Visual saliency estimation for 360° images using stacked autoencoder.☆14Jun 5, 2017Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The code for paper FLDCF, with various forgery detection and localization methods.☆19Mar 16, 2026Updated 6 months ago
- StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback☆73Aug 31, 2024Updated 2 years ago
- Official InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows☆20Nov 4, 2025Updated 11 months ago
- ☆47Jun 11, 2025Updated last year
- This is the official code for the paper "Lazy Safety Alignment for Large Language Models against Harmful Fine-tuning" (NeurIPS2024)☆28Sep 10, 2024Updated 2 years ago
- SalGan pytorch☆10Sep 3, 2018Updated 8 years ago
- WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs☆51Jul 12, 2026Updated 2 months ago