-
National Institute of Advanced Industrial Science and Technology
- https://kenshohara.github.io/
Stars
LAVIS - A One-stop Library for Language-Vision Intelligence
[ECCV 2022] Is Appearance Free Action Recognition Possible?
A header-only, single-file library for colormaps written in C++11
Accelerator of Scientific Development and Research. A project template developed by XCCV group of cvpaper.challenge.
PyTorch implementation of BEVT (CVPR 2022) https://arxiv.org/abs/2112.01529
[NeurIPS 2022 Spotlight] VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training
Instant neural graphics primitives: lightning fast NeRF and more
This repo includes the source code and dataset information for reproducing the results of our paper (https://arxiv.org/abs/2009.06435)
PyTorch implementation of MAE https//arxiv.org/abs/2111.06377
A Tool for Extracting and Embedding Road Scene-Graphs
Build and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
A high-performance Python-based I/O system for large (and small) deep learning problems, with strong support for PyTorch.
Scenic: A Jax Library for Computer Vision Research and Beyond
Social Fabric: Tubelet Compositions for Video Relation Detection
A python library for self-supervised learning on images.
Spatial-Temporal Transformer for Dynamic Scene Graph Generation, ICCV2021
A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch
PyTorch code for Vision Transformers training with the Self-Supervised learning method DINO
PyTorch implementation of SimSiam https//arxiv.org/abs/2011.10566
An efficient video loader for deep learning with smart shuffling that's super easy to digest
The official pytorch implementation of our paper "Is Space-Time Attention All You Need for Video Understanding?"
A deep learning library for video understanding research.
MMGeneration is a powerful toolkit for generative models, based on PyTorch and MMCV.
Implementation of ViViT: A Video Vision Transformer


