Multimodal short video classification task, integrating video, image, audio and text modes for short video classification
☆20Mar 12, 2020Updated 6 years ago
Alternatives and similar repositories for Multimodal-short-video-classification
Users that are interested in Multimodal-short-video-classification are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 500,000 multimodal short video data and baseline models. 50万条多模态短视频数据集和基线模型(TensorFlow2.0)。☆136Jul 23, 2019Updated 7 years ago
- Hand Gesture Controlled Tello Drone using Python and OpenCV 2021☆12Jun 6, 2022Updated 4 years ago
- Modulated Fusion using Transformer for Linguistic-Acoustic Emotion Recognition☆33Dec 4, 2020Updated 5 years ago
- ☆11May 18, 2022Updated 4 years ago
- Implementation of the paper "Real-Time Emotion Recognition via Attention Gated Hierarchical Memory Network" in AAAI-2020.☆30Sep 2, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Multimodal Fusion, Multimodal Sentiment Analysis☆25Jun 20, 2020Updated 6 years ago
- ☆12Jan 16, 2022Updated 4 years ago
- Autonomous Exploration of mobile robots in unknown environments using Deep Reinforcement learning☆15Oct 28, 2023Updated 2 years ago
- Multimodal classification solution for the SIGIR eCOM using Co-attention and transformer language models☆19Aug 17, 2020Updated 6 years ago
- ☆25Jun 3, 2020Updated 6 years ago
- [ACMMM 2020] Code release for "Learning Deep Multimodal Feature Representation with Asymmetric Multi-layer Fusion"☆28Aug 19, 2021Updated 5 years ago
- An implementation of the paper 'Using Deep Networks for Scientific Discovery in Physiological Signals'☆12Aug 24, 2020Updated 6 years ago
- 3D sMRI data classification using PyTorch.☆15Aug 26, 2019Updated 7 years ago
- Engaged in research to help improve to boost text sentiment analysis using facial features from video using machine learning.☆32Jan 12, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- It is the code for ASSISTment data mining competition 2017☆13May 28, 2018Updated 8 years ago
- Official PyTorch implementation of Multilogue-Net (Best paper runner-up at Challenge-HML @ ACL 2020)☆58Dec 8, 2022Updated 3 years ago
- docker部署springboot项目的demo☆18Sep 24, 2023Updated 3 years ago
- Converting VIS json label to VOS format☆12Feb 16, 2021Updated 5 years ago
- 多模态视频分类模型☆33Nov 23, 2022Updated 3 years ago
- Langchain_CrewAI_Gemini - An Gemini AI powered AI Agent (Multi-Agent) Project.☆14Mar 24, 2024Updated 2 years ago
- Jupyter Notebooks with Titanic Classification using Decision Trees and Random Forest☆14Aug 26, 2017Updated 9 years ago
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- ☆17Feb 18, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- AI Drone with embedded GPU (Nvidia Jetson Nano) for computer vision and autonomous flight☆20Nov 29, 2021Updated 4 years ago
- Lectures and Labs for Data Parallel Computing and DPC++. Sponsored by Intel Corporation.☆13Apr 7, 2022Updated 4 years ago
- This Pytorch repo uses BiConvLSTM in a Spatiotemporal Encoder to detect violence in Videos. Three benchmark datasets namely Hockey, Movie…☆38Sep 5, 2018Updated 8 years ago
- Large Language-and-Vision Assistant for BioMedicine, built towards multimodal GPT-4 level capabilities.☆10Nov 29, 2023Updated 2 years ago
- Reference implementation and test synthetic data for Sorted Center Time echo density measure for acoustic impulse responses☆15Mar 18, 2020Updated 6 years ago
- A PyTorch implementation of the paper Multimodal Transformer with Multiview Visual Representation for Image Captioning☆25Sep 4, 2020Updated 6 years ago
- Implement a GRU/LSTM model using Keras, and train it to classify the languages using MFCC features☆25Aug 2, 2024Updated 2 years ago
- ☆13Mar 25, 2021Updated 5 years ago
- Code and data for "Medical Dialogue Generation via Dual Flow Modeling" (ACL 2023 Findings)☆14Nov 22, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A smart autonomous drone with Object Tracking and Object Detection capabilities☆16Jul 3, 2022Updated 4 years ago
- Implement attention model to LSTM using TensorFlow☆10Jul 3, 2018Updated 8 years ago
- Submission to MediaEval 2021 Emotions and Themes in Music challenge. Noisy-student training for music emotion tagging☆11Dec 2, 2021Updated 4 years ago
- Deployed a facial emotion recognition using neural network model which predicts the emotion from faces in images, videos and live feed fr…☆12May 2, 2021Updated 5 years ago
- ☆10Nov 10, 2021Updated 4 years ago
- ☆25Apr 23, 2019Updated 7 years ago
- Acoustic Scene Classification using transfer learning on VGGish pre-trained model☆11Jan 3, 2018Updated 8 years ago