基于Qwen Agent框架,融合JAKA机械臂、视觉检测、语音识别与合成、MCP数据库的多模态大模型
☆24May 26, 2025Updated last year
Alternatives and similar repositories for MIRA-Multimodal-Intelligent-Robotic-Assistant
Users that are interested in MIRA-Multimodal-Intelligent-Robotic-Assistant are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 机械臂控制系统项目 - 包含传统PID控制和强化学习抓取系统,VLM+RL正在集成中......☆23Aug 31, 2025Updated last year
- PhysVLM: Enabling Visual Language Models to Understand Robotic Physical Reachability☆42Mar 18, 2025Updated last year
- 🦾 PyTorch Implementation for the ICRA'24 Paper, "PROGrasp: Pragmatic Human-Robot Communication for Object Grasping"☆15May 5, 2025Updated last year
- Official repo for the 2024 CoRL Paper: EXTRACT: Efficient Policy Learning by Extracting Transferable Robot Skills from Offline Data☆19Apr 21, 2025Updated last year
- Vision based RL agent to control a UR5 arm in Human-Robot collaborative environments☆31Jul 13, 2026Updated 2 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- 利用pytorch实现图像分类的一个完整的代码,训练,预测,TTA,模型融合,模型部署,cnn提取特征,svm或者随机森林等进行分类,模型蒸馏,一个完整的代码☆31Dec 15, 2020Updated 5 years ago
- Personal Experiment around ReKep☆18Feb 3, 2025Updated last year
- ☆21Jun 20, 2025Updated last year
- Task and Motion Planning with Uncertainty and Risk Awareness☆27May 14, 2025Updated last year
- scripted and traced deep learning models and their applications☆18Mar 21, 2023Updated 3 years ago
- An end-to-end demo of an autonomous agentic mobile manipulator for warehouse robotics, showcasing fully on-device perception, reasoning, …☆47Jul 15, 2026Updated 2 months ago
- ☆15Oct 22, 2024Updated last year
- This is the source code to paper “DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation”.☆36Aug 13, 2025Updated last year
- [IROS 2025] Official code of ”HybridTM: Combining Transformer and Mamba for 3D Semantic Segmentation“☆26Jul 25, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆17Oct 2, 2021Updated 4 years ago
- 通过手动标注数据集,训练老人摔倒的模型☆24Mar 27, 2023Updated 3 years ago
- ☆20Apr 1, 2026Updated 5 months ago
- OpenVLA Lightweight Version(0.5B). It uses qwen2-0.5B and fine-tunes using mllm format, without occupying LLM's inherent tokens. It repre…☆19Jan 7, 2026Updated 8 months ago
- Use yolov5 to realize the road occupation operation and vehicle parking violation detection in urban streets, and can independently delin…☆13Jan 2, 2023Updated 3 years ago
- DeepSeek controls the panda robotic arm☆53Nov 23, 2025Updated 9 months ago
- Strawberry detection by using the YOLOv8 model, Ripeness measurement of strawberry by using OpenCV and Measure distance of strawberry fro…☆13Oct 31, 2023Updated 2 years ago
- This AI-powered system automates attendance tracking in schools, colleges, and workplaces using face recognition. It replaces traditional…☆24May 13, 2025Updated last year
- A framework for integrated task and motion planning from perception☆33Dec 31, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆19Mar 4, 2026Updated 6 months ago
- This project is an AI-driven car parking monitoring system that uses YOLO (You Only Look Once) object detection and OpenCV to detect occu…☆19Feb 20, 2025Updated last year
- ☆15Apr 28, 2023Updated 3 years ago
- Code for FLIP: Flow-Centric Generative Planning for General-Purpose Manipulation Tasks☆85Dec 12, 2024Updated last year
- Code implementation of paper "MUSE: Mamba is Efficient Multi-scale Learner for Text-video Retrieval (AAAI2025)"☆26Feb 2, 2025Updated last year
- ☆24May 6, 2025Updated last year
- This repository features projects that track physical activities using computer vision, integrating OpenCV with MediaPipe's Pose Estimati…☆15Oct 21, 2024Updated last year
- The collections of MOE (Mixture Of Expert) papers, code and tools, etc.☆12Mar 15, 2024Updated 2 years ago
- 这个仓库是使用Yolov8-Seg实例分割算法和Sgbm深度估计算法结合的例子,可以作为智能小车,无人机,水下航行器导航的视觉识别部分,为基于深度学习的避障导航作为参考。☆15Jun 30, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This project implements a real-time customer tracking system using computer vision to monitor how long customers spend in a store. It com…☆11Dec 27, 2024Updated last year
- 学习电机控制算法的工程,持续更新中。。。☆20Aug 15, 2024Updated 2 years ago
- 机械臂定位抓取☆207Dec 26, 2023Updated 2 years ago
- [ICCV2025] CleanPose: Category-Level Object Pose Estimation via Causal Learning and Knowledge Distillation☆26Sep 16, 2025Updated last year
- 在监控画质下实现对校园自行车的重识别,包含REID模型识别,向量数据库检索,UI展示☆11Feb 13, 2024Updated 2 years ago
- A Hand Gesture Volume Control application made using OpenCV & Mediapipe☆11Jul 29, 2022Updated 4 years ago
- YOLO v8과 SAM (Sagment Anything Model)을 결합한 해충 (pest) detection model☆13Apr 28, 2024Updated 2 years ago