Labels
Labels
RFC: Under Review
Story
Story: ATen Op Support
Story: Binding Names
Story: Build & Install & Packaging
Installation failures, build system issues, Jetson/aarch64 support, Windows packaging, and dependenstory: Build Install and Packaging
story: CI/CD & Testing Infrastructure
Test suite failures across NGC/DLFW/Orin/Windows platforms, CI pipeline improvements, and automatedstory: Documentation & Examples
Broken or outdated examples, missing guides, user questions about usage patterns, and documentationstory: Dynamic Shapes & Symbolic Tracing
Dynamic input shapes, symbolic shape inference failures, opt/min/max shape handling, and data-dependStory: Dynamo Compile Improvements
story: Dynamo Frontend & Partitioning
torch.compile, torch.export, FX graph tracing, graph partitioner, graph breaks, and the Dynamo-to-TRStory: Export/Compile Unification
Story: Infrastructure Upgrades
story: LLM & Generative AI
Large language models (GPT2, Llama, Mistral, Qwen), diffusion models (FLUX, SD), VLMs, MoE, attentioStory: Multi-GPU & Distributed & Custom Ops
Multi-GPU inference, NCCL optimization, distributed training integration, custom CUDA kernels, andstory: Multi-GPU Distributed and Custom Ops
story: Operator Coverage & Converters
Missing ATen op converters, converter accuracy bugs, type-casting issues, and new operator support rstory: Performance & Benchmarking
Performance gaps vs ONNX-TRT, engine size, reformatting overhead, profiling tools, and throughput reStory: Porting CI/CD to GHA
story: Quantization & Precision
INT8/FP8/QAT quantization, enabled_precisions deprecation, AutoCast, weak typing, and low-precisionStory: Runtime & Memory & Serialization
Engine execution, CUDA graphs, weight streaming, model save/load, refit, memory management, and runstory: Runtime and Memory
Story: TensorRT 8
Story: Transformer Encoder
story: TRT-RTX Platform
Features and bugs specific to the TensorRT-RTX backend, Thor platform, and RTX-optimized execution ptorchscript
Upstreaming PR
WAR
WIP
wontfix