Parameter-efficient finetuning script for Phi-3-vision, the strong multimodal language model by Microsoft.
☆58Jun 17, 2024Updated 2 years ago
Alternatives and similar repositories for Phi3V-Finetuning
Users that are interested in Phi3V-Finetuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A bug-free and improved implementation of LLaVA-UHD, based on the code from the official repo☆35Aug 12, 2024Updated 2 years ago
- ☆17Jun 9, 2024Updated 2 years ago
- Implementation of ICCV 2025 paper "Growing a Twig to Accelerate Large Vision-Language Models".☆31May 23, 2026Updated 4 months ago
- ☆85Feb 1, 2024Updated 2 years ago
- Quick exploration into fine tuning florence 2☆340Sep 19, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- a family of highly capabale yet efficient large multimodal models☆194Aug 23, 2024Updated 2 years ago
- A reproduction of RetinaFace by PaddlePaddle☆14Dec 19, 2021Updated 4 years ago
- A library for simplifying training with multi gpu setups in the HuggingFace / PyTorch ecosystem.☆16Jun 10, 2026Updated 3 months ago
- ☆13Jul 10, 2024Updated 2 years ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 5 months ago
- ICCV 2025 Code for "Salvaging the Overlooked: Leveraging Class-Aware Contrastive Learning for Multi-Class Anomaly Detection"☆11Nov 26, 2025Updated 10 months ago
- This is a Phi Family of SLMs book for getting started with Phi Models. Phi a family of open sourced AI models developed by Microsoft. Phi…☆3,884Sep 9, 2026Updated 2 weeks ago
- Retrieval-augmented Image Captioning☆13Feb 16, 2023Updated 3 years ago
- Embroid: Unsupervised Prediction Smoothing Can Improve Few-Shot Classification☆11Aug 12, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICLR 2025] Official implementation of paper "Dynamic Low-Rank Sparse Adaptation for Large Language Models".☆25Mar 16, 2025Updated last year
- Imagen-mini for girl image generation☆12Nov 19, 2022Updated 3 years ago
- Diapositivas, notebooks y material de charlas, talleres y el grupo de estudio☆12Apr 24, 2024Updated 2 years ago
- FastAPI WebSocket server for the OpenVoice text-to-speech model.☆12Jun 6, 2024Updated 2 years ago
- This repository demonstrates the data preparation and fine-tuning the IDEFICS Vision Language Model.☆26May 16, 2024Updated 2 years ago
- Infusion: Preventing Customized Text-to-Image Diffusion from Overfitting☆14Dec 19, 2025Updated 9 months ago
- Build your own custom knowledge base from various sources such as youtube videos transcripts, tweets, articles, videos and audios. Uses G…☆13Dec 15, 2023Updated 2 years ago
- Pipeline to scrape prompt + image url pairs from LAION `share-dalle-3` discord channel☆11Oct 10, 2023Updated 2 years ago
- A Web app demonstrating multimodal image search using Visualized-BGE model☆16Dec 1, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆87Nov 14, 2023Updated 2 years ago
- ☆16Dec 11, 2023Updated 2 years ago
- The system of SUDA-HUAWEI submitted at CAMR2022.☆12Nov 22, 2022Updated 3 years ago
- CVPR 2023: PAniC-3D, rendering☆16Mar 25, 2023Updated 3 years ago
- ☆13Apr 3, 2026Updated 5 months ago
- ☆14Feb 7, 2024Updated 2 years ago
- MicroPython viper documentation and examples☆16Apr 19, 2024Updated 2 years ago
- Increase the inference speed of the model☆19Jun 7, 2022Updated 4 years ago
- An NVIDIA Triton Server workflow for OCR and the LayoutLMv3 Transformer Model☆30Sep 14, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Monitoring of a GPU system sending either Slack or Mattermost messages via webhooks☆12Jul 20, 2017Updated 9 years ago
- Multi-lingual AudioCaps☆14Nov 20, 2023Updated 2 years ago
- Tutorial for text classification with BERT, using HuggingFace's transformers.☆13Jan 15, 2020Updated 6 years ago
- Retrieve XPath and CSS selectors from elements selected in Playwright☆16Jun 17, 2022Updated 4 years ago
- CCL2024 Chinese Essay Rhetoric Recognition and Understanding☆17Oct 1, 2024Updated last year
- ☆14May 9, 2024Updated 2 years ago
- Towards Real-World Writing Assistance: A Chinese Character Checking Benchmark with Faked and Misspelled Characters☆17May 30, 2024Updated 2 years ago