Tiny Kinetics-400 for test

Last update: Jan 06, 2023

Related tags

Deep Learning tiny-kinetics-400

Overview

Kinetics-400迷你数据集

English | 简体中文

该数据集旨在解决的问题：参照Kinetics-400数据格式，训练基于自己数据的视频理解模型。

数据集介绍

Kinetics-400是视频领域benchmark常用数据集，详细介绍可以参考其官方网站Kinetics。整个数据集包含400个类别，全部文件大概需要135G左右的存储空间，下载起来比较困难。

Tiny-Kinetics-400同样包含400个类别，每个类别下仅有两条视频数据，分为train与val，可用于调试一些视频理解模型。

具体对比如下：

数据集	训练条数	验证条数	大小
Kinetics-400	234619	19761	135G
Tiny-Kinetics-400	400	400	420M

Tiny-Kinetics-400下载

目前提供了百度网盘的下载方式：

下载方式	链接
百度云	BaiduCloud (1cns)

抽帧Extract Frames

通常在训练视频理解模型时，会提前对视频文件进行抽帧，以此来加速训练过程。这里提供了抽帧脚本，且满足以下条件：

每个视频只抽取300帧
如果整个视频多于300帧，直接舍弃之后的视频帧
如果整个视频少于300帧，复制最后的视频帧以填充至300帧

使用方式：

python ./tools/extract_frames.py --source_dir ~/data/tiny-kinetics-400/train_256 ~/data/kinetics400_30fps_frames/train
python ./tools/extract_frames.py --source_dir ~/data/tiny-kinetics-400/val_256 ~/data/kinetics400_30fps_frames/val

将meta文件移到视频帧目录下：

mv ./annotations/tiny_train.csv ~/data/kinetics400_30fps_frames/
mv ./annotations/tiny_val.csv ~/data/kinetics400_30fps_frames/

最终的目录结构如下：

kinetics400_30fps_frames/
├── train/
│   ├── abseiling/
│   │   ├──_4YTwq0-73Y_000044_000054
│   │   │  ├──frame_00001.jpg
│   │   │  ├──...
│   │   ├──...
│   ├──...
├── val/
│   ├── abseiling/
│   │   ├──-3B32lodo2M_000059_000069
│   │   │  ├──frame_00001.jpg
│   │   │  ├──...
│   │   ├──...
│   ├──...
├── tiny_train.csv
├── tiny_val.csv

TODO

更多下载方式

Tiny Kinetics-400 for test

Related tags

Overview

Kinetics-400迷你数据集

数据集介绍

Tiny-Kinetics-400下载

抽帧Extract Frames

TODO

参考

Owner

A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation

Locally Constrained Self-Attentive Sequential Recommendation

Optimizing DR with hard negatives and achieving SOTA first-stage retrieval performance on TREC DL Track (SIGIR 2021 Full Paper).

A two-stage U-Net for high-fidelity denoising of historical recordings

exponential adaptive pooling for PyTorch

Compare outputs between layers written in Tensorflow and layers written in Pytorch

The source code for Generating Training Data with Language Models: Towards Zero-Shot Language Understanding.

Official project website for the CVPR 2021 paper "Exploring intermediate representation for monocular vehicle pose estimation"

MetaDrive: Composing Diverse Scenarios for Generalizable Reinforcement Learning

Tacotron 2 - PyTorch implementation with faster-than-realtime inference

PyTorch Implementation of "Non-Autoregressive Neural Machine Translation"

CLIPort: What and Where Pathways for Robotic Manipulation

Model that predicts the probability of a Twitter user being anti-vaccination.

Portfolio asset allocation strategies: from Markowitz to RNNs

Implementation of the paper NAST: Non-Autoregressive Spatial-Temporal Transformer for Time Series Forecasting.

This is the official pytorch implementation for the paper: Instance Similarity Learning for Unsupervised Feature Representation.

Code for the paper Hybrid Spectrogram and Waveform Source Separation

Repository of Vision Transformer with Deformable Attention

TLoL (Python Module) - League of Legends Deep Learning AI (Research and Development)

hySLAM is a hybrid SLAM/SfM system designed for mapping