This repository has been archived on 2026-02-22. You can view files and clone it. You cannot open issues or pull requests or push a commit.
2017-08-10 19:21:15 +02:00
2017-08-10 19:21:15 +02:00
2017-08-10 19:21:15 +02:00
2017-08-10 16:22:17 +02:00
2017-08-03 16:09:23 +02:00
2017-08-10 16:22:17 +02:00
2017-08-09 14:13:17 +00:00

tkDNN

tkDNN is a Deep Neural Network library built with cuDNN primitives specifically thought to work on NVIDIA TK1 board.
The main scope is to do high performance inference on already trained models.

this branch is actually work on every NVIDIA GPU that support the dependencies:

  • CUDA 8
  • CUDNN 6
  • TENSORRT 2

Workflow

The recommended workflow follow these step:

  • Build and train a model in Keras (on any PC)
  • Export weights and bias
  • Define the model on tkDNN
  • Do inference (on TK1)

Compile the library

Build with cmake

mkdir build
cd build
cmake ..
make

during the cmake configuration it will be dowloaded the weights needed for running the tests

Test

Assumiung you have correctly builded the library these are the test ready to exec:

  • test_simple: a simple convolutional and dense network (CUDNN only)
  • test_mnist: the famous mnist netwok (CUDNN and TENSORRT)
  • test_mnistRT: the mnist network hardcoded in using tensorRT apis (TENSORRT only)
  • test_yolo: YOLO detection network (CUDNN and TENSORRT)
  • test_yolo_tiny: smaller version of YOLO (CUDNN and TENSRRT)
S
Description
Deep neural network library and toolkit to do high performace inference on NVIDIA jetson platforms [MIRROR]
Readme GPL-2.0 74 MiB
version 0.6 Latest
2021-07-23 14:37:04 +02:00
Languages
C++ 90.9%
Cuda 4%
Python 2.3%
Shell 1.1%
CMake 1.1%
Other 0.6%