NVIDIA Releases TensorRT Model Connect in Public Preview
Tool converts supported Hugging Face models to TensorRT inference in two commands without ONNX export.
NVIDIA AI posted that the company released TensorRT Model Connect in public preview. The account stated users can take a supported Hugging Face model to end-to-end TensorRT inference in two commands with no intermediate ONNX export and that the resulting bundle runs through native C++ APIs. NVIDIA also said the project was built with OpenAI Codex agents with humans directing and reviewing the work. Tianqi Chen separately noted first-class support for TVM-FFI in the release, enabling custom DSL and agent-generated kernels inside TensorRT inference.
Combined views
96.2K
2 posts, first seen 11d ago