Optimizing NMT with TensorRT
OpenNMT is an open source neural machine translation and neural machine sequencing model. Using Volta Tensor Cores and TensorRT, we''re able to improve performance by 100 times over CPU implementation. We''ll discuss OpenNMT and how we implement it via TensorRT. We''ll show how by using our plugin interface and new TensorRT features, we''re able to implement this network at high performance.
OpenNMT is an open source neural machine translation and neural machine sequencing model. Using Volta Tensor Cores and TensorRT, we''re able to improve performance by 100 times over CPU implementation. We''ll discuss OpenNMT and how we implement it via TensorRT. We''ll show how by using our plugin interface and new TensorRT features, we''re able to implement this network at high performance.
Back
Keywords:
AI Application Deployment and Inference, Advanced AI Learning Techniques (incl. GANs and NTMs), GTC Silicon Valley 2018 - ID S8822