tvm-vta

Deep Learning Accelerator

A comprehensive hardware design stack for accelerating deep learning models

Open, Modular, Deep Learning Accelerator

GitHub

258 stars
40 watching
73 forks
Language: Scala
last commit: over 2 years ago
Linked from 1 awesome list

hardwaremachine-learningtensortvmvta

Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
nvdla/hwThe NVDLA project provides hardware designs and tools for building deep learning inference accelerators.1,763
vlang/vtlA C library providing an n-dimensional tensor data structure and linear algebra routines148
doonny/pipecnnA tool for accelerating convolutional neural networks on Field-Programmable Gate Arrays (FPGAs) using OpenCL-based hardware design1,264
homles11/igcv3An implementation of an efficient deep neural network architecture189
eaplatanios/tensorflow_scalaA Scala API for TensorFlow's deep learning functionality939
vict0rsch/deep_learningA collection of tutorials and resources on implementing deep learning models using Python libraries such as Keras and Lasagne.426
jnhwkim/nips-mrn-vqaThis project presents a neural network model designed to answer visual questions by combining question and image features in a residual learning framework.39
acceleratehs/accelerate-llvmCompiles Accelerate code to LLVM IR and executes it on CPUs or NVIDIA GPUs159
google/cfu-playgroundA framework for designing and evaluating custom processor instructions to accelerate machine learning tasks on FPGAs.476
coreylowman/dfdxA deep learning library for Rust with GPU acceleration and ergonomic API.1,754
vlgiitr/dmn-plusA PyTorch implementation of an improved question answering architecture with dynamic memory networks and attention mechanisms64
vlfeat/autonnAn API wrapper around MatConvNet that adds automatic differentiation for easy deep learning prototyping and research89
uber/petastormEnables training and evaluation of deep learning models from Apache Parquet datasets in various machine learning frameworks1,805
intel/intel-extension-for-tensorflowEnables heterogeneous high-performance computing on Intel CPUs and GPUs for deep learning workloads323
mit-han-lab/proxylessnasDirect neural architecture search on target task and hardware for efficient model deployment1,429