tensorRT配合triton部署模型

文章目录

一、onnx

1)onnx格式介绍

2)onnx模型网络图认识

initializer:

拓扑关系:先conv,后relu

3)onnx关键数据结构(边+算子=》组成图=》组成模型)

3.1 边

3.2 算子

3.3 模型

3.4 图

4)onnx原生API搭建onnx模型

  • 指定节点

    ①resize节点

    ②conv节点

    ③Add节点

  • 步骤

    ①定义tensor节点,定义输入、输出

    ②制作节点

    ③根据节点制作图和模型

    ④保存成onnx

  • 代码

python 复制代码
import onnx
from onnx import helper
from onnx import TensorProto
import onnxruntime
import numpy as np
# define tensor
input = helper.make_tensor_value_info('input', TensorProto.FLOAT, [1,3,256, 256])
roi = helper.make_tensor_value_info('roi', TensorProto.FLOAT, [])
scales = helper.make_tensor_value_info('scales', TensorProto.FLOAT, [4])
conv_input = helper.make_tensor_value_info('conv_input', TensorProto.FLOAT, [1,3,512,512])
conv_weight = helper.make_tensor_value_info('conv_weight', TensorProto.FLOAT, [32,3,3,3])
conv_bias = helper.make_tensor_value_info('conv_bias', TensorProto.FLOAT, [32])
conv_output = helper.make_tensor_value_info('conv_output', TensorProto.FLOAT, [1,32,512,512])
add_input = helper.make_tensor_value_info('add_input', TensorProto.FLOAT, [1])
output = helper.make_tensor_value_info('output', TensorProto.FLOAT, [1,32,512,512])

# make node
resize_node = helper.make_node("Resize", ['input','roi','scales'], ['conv_input'], name='resize')
conv_node = helper.make_node("Conv", ['conv_input','conv_weight','conv_bias'], ['conv_output'], name='conv',strides=[1, 1],pads=[1, 1, 1, 1])
add_node = helper.make_node('Add', ['conv_output','add_input'], ['output'], name='add')

# make graph
graph = helper.make_graph([resize_node,conv_node,add_node],'resize_conv_add_graph',inputs=[input,roi,scales,conv_weight,conv_bias,add_input],outputs=[output])

# make model
model = helper.make_model(graph, opset_imports=[helper.make_opsetid('', 21)]) # 构建模型
onnx.checker.check_model(model)  # 检测模型的准确性
  • 输出的模型结构
shell 复制代码
ir_version: 11
graph {
  node {
    input: "input"
    input: "roi"
    input: "scales"
    output: "conv_input"
    name: "resize"
    op_type: "Resize"
  }
  node {
    input: "conv_input"
    input: "conv_weight"
    input: "conv_bias"
    output: "conv_output"
    name: "conv"
    op_type: "Conv"
    attribute {
      name: "pads"
      ints: 1
      ints: 1
      ints: 1
      ints: 1
      type: INTS
    }
    attribute {
      name: "strides"
      ints: 1
      ints: 1
      type: INTS
    }
  }
  node {
    input: "conv_output"
    input: "add_input"
    output: "output"
    name: "add"
    op_type: "Add"
  }
  name: "resize_conv_add_graph"
  input {
    name: "input"
    type {
      tensor_type {
        elem_type: 1
        shape {
          dim {
            dim_value: 1
          }
          dim {
            dim_value: 3
          }
          dim {
            dim_value: 256
          }
          dim {
            dim_value: 256
          }
        }
      }
    }
  }
  input {
    name: "roi"
    type {
      tensor_type {
        elem_type: 1
        shape {
        }
      }
    }
  }
  input {
    name: "scales"
    type {
      tensor_type {
        elem_type: 1
        shape {
          dim {
            dim_value: 4
          }
        }
      }
    }
  }
  input {
    name: "conv_weight"
    type {
      tensor_type {
        elem_type: 1
        shape {
          dim {
            dim_value: 32
          }
          dim {
            dim_value: 3
          }
          dim {
            dim_value: 3
          }
          dim {
            dim_value: 3
          }
        }
      }
    }
  }
  input {
    name: "conv_bias"
    type {
      tensor_type {
        elem_type: 1
        shape {
          dim {
            dim_value: 32
          }
        }
      }
    }
  }
  input {
    name: "add_input"
    type {
      tensor_type {
        elem_type: 1
        shape {
          dim {
            dim_value: 1
          }
        }
      }
    }
  }
  output {
    name: "output"
    type {
      tensor_type {
        elem_type: 1
        shape {
          dim {
            dim_value: 1
          }
          dim {
            dim_value: 32
          }
          dim {
            dim_value: 512
          }
          dim {
            dim_value: 512
          }
        }
      }
    }
  }
}
opset_import {
  domain: ""
  version: 21
}

5)onnx模型推理

6)dump模型,输出onnx各算子信息

  • 代码
python 复制代码
"""
打印onnx节点信息
"""
import onnx
import onnxruntime as rt
import numpy as np
 
# 加载ONNX模型
model_path = 'resize_conv_add.onnx'
onnx_model = onnx.load(model_path)
session = rt.InferenceSession(model_path) #类似于tf.Session
input_name = session.get_inputs()[0].name
roi_name = session.get_inputs()[1].name
scales_name = session.get_inputs()[2].name
conv_weight_name = session.get_inputs()[3].name
conv_bias_name = session.get_inputs()[4].name
add_input_name = session.get_inputs()[5].name
output_name = session.get_outputs()[0].name
intermediate_layer_names = [onnx_model.graph.node[i].name for i in range(len(onnx_model.graph.node))]
print(f"input_name:{input_name}, conv_weight_name: {conv_weight_name}")
print('node=',onnx_model.graph.node)
for node in onnx_model.graph.node:
    print('node_name=',node.name)
    print('node_input=',node.input)
    print('node_output=',node.output)

7)onnx模型实用工具: onnx graphsurgeon

8)onnx模型实用工具: onnx simplier

9)onnx与TensorRT模型部署的前后纠葛

相关推荐
BestSongC9 小时前
基于VUE和FastAPI的行人目标检测系统
vue.js·人工智能·yolo·目标检测·fastapi
王哈哈^_^20 小时前
YOLO11实例分割训练任务——从构建数据集到训练的完整教程
人工智能·深度学习·算法·yolo·目标检测·机器学习·计算机视觉
xuehaikj1 天前
香烟品牌识别与分类:yolov5-LSKNet模型应用
yolo·数据挖掘
Sunhen_Qiletian1 天前
YOLO的再进步---YOLOv3算法详解(上)
算法·yolo·计算机视觉
王哈哈^_^1 天前
【完整源码+数据集】中药材数据集,yolov8中药分类检测数据集 9709 张,中药材分类识别数据集,中药材识别系统实战教程
人工智能·深度学习·算法·yolo·目标检测·计算机视觉·毕业设计
深度学习lover1 天前
<数据集>yolo遥感航拍船舶识别数据集<目标检测>
人工智能·python·yolo·目标检测·计算机视觉·航拍船舶识别
深蓝海拓1 天前
YOLO v11的学习记录(六) 把标注好的大图切割成小图
深度学习·学习·yolo
minhuan2 天前
构建AI智能体:九十五、YOLO视觉大模型入门指南:从零开始掌握目标检测
人工智能·yolo·目标检测·计算机视觉·视觉大模型
xuehaikj2 天前
YOLOv8多场景人物识别定位与改进ASF-DySample算法详解
算法·yolo·目标跟踪
番茄小能手2 天前
【全网唯一】按键精灵安卓版纯本地离线yolo插件
yolo