Tiny Kinetics-400 for test

Last update: Jan 06, 2023

Related tags

Deep Learning tiny-kinetics-400

Overview

Kinetics-400迷你数据集

English | 简体中文

该数据集旨在解决的问题：参照Kinetics-400数据格式，训练基于自己数据的视频理解模型。

数据集介绍

Kinetics-400是视频领域benchmark常用数据集，详细介绍可以参考其官方网站Kinetics。整个数据集包含400个类别，全部文件大概需要135G左右的存储空间，下载起来比较困难。

Tiny-Kinetics-400同样包含400个类别，每个类别下仅有两条视频数据，分为train与val，可用于调试一些视频理解模型。

具体对比如下：

数据集	训练条数	验证条数	大小
Kinetics-400	234619	19761	135G
Tiny-Kinetics-400	400	400	420M

Tiny-Kinetics-400下载

目前提供了百度网盘的下载方式：

下载方式	链接
百度云	BaiduCloud (1cns)

抽帧Extract Frames

通常在训练视频理解模型时，会提前对视频文件进行抽帧，以此来加速训练过程。这里提供了抽帧脚本，且满足以下条件：

每个视频只抽取300帧
如果整个视频多于300帧，直接舍弃之后的视频帧
如果整个视频少于300帧，复制最后的视频帧以填充至300帧

使用方式：

python ./tools/extract_frames.py --source_dir ~/data/tiny-kinetics-400/train_256 ~/data/kinetics400_30fps_frames/train
python ./tools/extract_frames.py --source_dir ~/data/tiny-kinetics-400/val_256 ~/data/kinetics400_30fps_frames/val

将meta文件移到视频帧目录下：

mv ./annotations/tiny_train.csv ~/data/kinetics400_30fps_frames/
mv ./annotations/tiny_val.csv ~/data/kinetics400_30fps_frames/

最终的目录结构如下：

kinetics400_30fps_frames/
├── train/
│   ├── abseiling/
│   │   ├──_4YTwq0-73Y_000044_000054
│   │   │  ├──frame_00001.jpg
│   │   │  ├──...
│   │   ├──...
│   ├──...
├── val/
│   ├── abseiling/
│   │   ├──-3B32lodo2M_000059_000069
│   │   │  ├──frame_00001.jpg
│   │   │  ├──...
│   │   ├──...
│   ├──...
├── tiny_train.csv
├── tiny_val.csv

TODO

更多下载方式

Tiny Kinetics-400 for test

Related tags

Overview

Kinetics-400迷你数据集

数据集介绍

Tiny-Kinetics-400下载

抽帧Extract Frames

TODO

参考

Owner

Generic U-Net Tensorflow implementation for image segmentation

PoseViz – Multi-person, multi-camera 3D human pose visualization tool built using Mayavi.

More Photos are All You Need: Semi-Supervised Learning for Fine-Grained Sketch Based Image Retrieval

A python library for face detection and features extraction based on mediapipe library

Certified Patch Robustness via Smoothed Vision Transformers

Notification Triggers for Python

Neural Scene Flow Fields for Space-Time View Synthesis of Dynamic Scenes

Replication attempt for the Protein Folding Model

Similarity-based Gray-box Adversarial Attack Against Deep Face Recognition

This is Unofficial Repo. Lips Don't Lie: A Generalisable and Robust Approach to Face Forgery Detection (CVPR 2021)

Official Pytorch implementation of "Learning Debiased Representation via Disentangled Feature Augmentation (Neurips 2021, Oral)"

Code for paper "Document-Level Argument Extraction by Conditional Generation". NAACL 21'

Using Machine Learning to Test Causal Hypotheses in Conjoint Analysis

This repo contains the implementation of YOLOv2 in Keras with Tensorflow backend.

This repository contains part of the code used to make the images visible in the article "How does an AI Imagine the Universe?" published on Towards Data Science.

DCGAN-tensorflow - A tensorflow implementation of Deep Convolutional Generative Adversarial Networks

[ICCV 2021] Released code for Causal Attention for Unbiased Visual Recognition

LSTMs (Long Short Term Memory) RNN for prediction of price trends

The `rtdl` library + The official implementation of the paper

LSSY量化交易系统