812 skills found · Page 4 of 28
qijiezhao / Video Classification Action Recognitionsome famous works and new content to be added
sujiongming / Awesome Video Understandingvideo-understanding:Video Classification, Action Recognition, Video Datasets
ahmetozlu / Face Recognition CropMulti-view face recognition, face cropping and saving the cropped faces as new images on videos to create a multi-view face recognition database.
whwu95 / MVFNet【AAAI'2021】MVFNet: Multi-View Fusion Network for Efficient Video Recognition
danduncan / HappyNetConvolutional neural network that does real-time emotion recognition. HappyNet detects faces in video and images, classifies the emotion on each face, then replaces each face with the correct emoji for that emotion. Based on Caffe and the "Emotions in the Wild" network available on Caffe model zoo.
blackfeather-wang / AdaFocusReducing spatial redundancy in video recognition. SOTA computational efficiency.
MSA-LMC / S2D[TAFFC 2024] The official implementation of paper: From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Recognition in Videos
wushidonguc / Two Stream Action Recognition KerasTwo-stream CNNs for video action recognition implemented in Keras
AlexanderMelde / SPHAR DatasetSurveillance Perspective Human Action Recognition Dataset: 7759 Videos from 14 Action Classes, aggregated from multiple sources, all cropped spatio-temporally and filmed from a surveillance-camera like position.
kenshohara / 3D ResNets3D ResNets for Action Recognition
ldkong1205 / TranSVAE[NeurIPS 2023] Unsupervised Video Domain Adaptation for Action Recognition: A Disentanglement Perspective
AkagawaTsurunaki / Zerolan CoreZerolanCore integrates many open-source, locally deployable AI models, and aims to integrate a series of AI models such as large language model (LLM), automatic speech recognition (ASR), text-to-speech (TTS), image captioning, optical character recognition (OCR), video captioning, etc.
lucazanella / AnomalyCLIPOfficial implementation of "Delving into CLIP latent space for Video Anomaly Recognition", CVIU 2024
rohitgirdhar / CATERCATER: A diagnostic dataset for Compositional Actions and TEmporal Reasoning
microsoft / ORBIT DatasetThe ORBIT dataset is a collection of videos of objects in clean and cluttered scenes recorded by people who are blind/low-vision on a mobile phone. The dataset is presented with a teachable object recognition benchmark task which aims to drive few-shot learning on challenging real-world data.
John1liu / YOLOV5 DeepSORT Vehicle Tracking MasterIn this project, urban traffic videos are collected from the middle section of Xi 'an South Second Ring Road with a large traffic flow, and interval frames are extracted from the videos to produce data sets for training and verification of YOLO V5 neural network. Combined with the detection results, the open-source vehicle depth model data set is used to train the vehicle depth feature weight file, and the deep-sort algorithm is used to complete the target tracking, which can realize real-time and relatively accurate multi-target recognition and tracking of moving vehicles.
Ha0Tang / HandGestureRecognition[Neurocomputing 2019] Fast and Robust Dynamic Hand Gesture Recognition via Key Frames Extraction and Feature Fusion
naver-ai / Tc Clip[ECCV 2024] Official PyTorch implementation of TC-CLIP "Leveraging Temporal Contextualization for Video Action Recognition"
TalalWasim / Video FocalNetsOfficial repository for "Video-FocalNets: Spatio-Temporal Focal Modulation for Video Action Recognition" [ICCV 2023]
Navu4 / Facial Recognition For Crime DetectionFace recognition software to detect criminals in images and videos, noting their time of occurences.