Where the RX 7800 XT stands today
trellis.cpp — ROCm build
A prebuilt ROCm binary avoids compiling gfx1101 kernels and avoids the Python dependency chain that has the most reported breakage. A native C++/GGML reimplementation of the whole TRELLIS.2 pipeline. Prebuilt ROCm archives are published for Linux and Windows, which removes the CUDA-only Python wheels from the problem entirely.
Open that guide →16 GB GDDR6. The upstream TRELLIS.2 project documents Linux and a 24 GB NVIDIA card as its standard path, so every route below is community work rather than a supported configuration. What follows separates what has actually been reported on this card from what is inferred from a shared instruction set.
What is actually reported
- A ROCm port of first-generation TRELLIS reports the RX 7800 XT as its tested card on ROCm 7.2.1 with a 16 GB minimum.
- The TRELLIS.2 ROCm ComfyUI guide names the 7800 XT among cards expected to work on the same Navi 31/32 patch set.
Memory and weight formats
Weight size is not peak VRAM. The figures below describe how much of the frame buffer the weights themselves occupy; the texture stage and working buffers sit on top of that. With 16 GB, these are the formats worth trying in order.
- Q8 · ≈ 9.5 GB · 16 GB practical
- Q4 · ≈ 6 GB · 12 GB practical, 8 GB tight
- GGUF K-quants (Q4_K_M – Q6_K) · ≈ 6 GB – 8 GB · 8 GB reported workable
Full detail on the two quantized options lives in the Q8 guide and the Q4 guide. If your workflow goes through ComfyUI rather than a native binary, the GGUF guide covers loader compatibility, which is the part that actually decides whether a quantized file opens.
What to watch for on this card
- The reported 7800 XT result is for the original TRELLIS, not TRELLIS.2. Do not read it as a TRELLIS.2 guarantee.
- That first-generation port also documents real fidelity trade-offs on RDNA 3, including silently culled triangles from a rasterizer bounds fix.
- 16 GB rules out comfortable f16 textured runs; plan for Q8 or Q4.
RDNA 3 in the full matrix
A separate install-guide project reports full shape and textured pipelines on a 7900 XTX under Linux. Navi 32 and Navi 33 cards share the family but are inferred, not individually reported.
| GPU | VRAM | ISA | ROCm forkLinux | ComfyUI LinuxLinux | ComfyUI WindowsWindows 11 | trellis.cpp ROCmLinux · Windows | trellis.cpp VulkanLinux · Windows |
|---|---|---|---|---|---|---|---|
| RDNA 3 gfx1100 · gfx1101 · gfx1102 | |||||||
| RX 7900 XTX | 24 GB | gfx1100 | ○Expected | ●Tested | –Unknown | ○Expected | ○Expected |
| RX 7900 XT | 20 GB | gfx1100 | ○Expected | ○Expected | –Unknown | ○Expected | ○Expected |
| RX 7800 XT | 16 GB | gfx1101 | ○Expected | ○Expected | –Unknown | ○Expected | ○Expected |
| RX 7700 XT | 12 GB | gfx1101 | –Unknown | ○Expected | –Unknown | ○Expected | ○Expected |
| RX 7600 XT | 16 GB | gfx1102 | –Unknown | ○Expected | –Unknown | ○Expected | ○Expected |
- ●Tested
A named project or maintainer reports this exact card completing the pipeline end to end.
- ◐Reported
Community reports exist but come from work in progress, partial runs, or a single tester.
- ○Expected
Inferred from a shared instruction set with a tested card. No direct report yet.
- ×Blocked
A known dependency, kernel, or runtime gap prevents this path today.
- –Unknown
No usable report either way. Treat as untested rather than as a failure.
Every AMD route on this page is community or third-party work. Neither Microsoft nor AMD lists TRELLIS.2 as an officially supported workload, and AMD's own documentation still limits Windows to PyTorch rather than the full ROCm stack.
TRELLIS.2 ROCm source fork
A fork of the upstream repository that ships HIP builds of FlexGEMM, CuMesh, nvdiffrast and nvdiffrec, plus a setup script that detects CUDA or ROCm and installs the matching dependencies. Validated on an RX 9070 XT 16 GB.
- OS
- Linux
- Stack
- ROCm 7.2 · PyTorch rocm7.2 · Python 3.10+
- Python runtime
- Required
ComfyUI + ROCm on Linux
The community ComfyUI wrapper plus a patch set that fixes the hardcoded GPU architecture flag, package paths and checkpoint downloads. Reported working end to end on a 7900 XTX for both shape-only and textured runs.
- OS
- Linux
- Stack
- ROCm 7.2.2 · PyTorch 2.11+rocm7.2 · Python 3.10–3.12
- Python runtime
- Required
ComfyUI + ROCm on Windows
Work in progress. RDNA 4 is confirmed running and testers are being recruited; RDNA 3 and RDNA 3.5 are next in line. The maintainer warns about silent bugs that can change the final output without raising an error.
- OS
- Windows 11
- Stack
- PyTorch 2.9.1 + ROCm 7.2.1, or PyTorch 2.12 + ROCm 7.14 · Python 3.12
- Python runtime
- Required
trellis.cpp — ROCm build
A native C++/GGML reimplementation of the whole TRELLIS.2 pipeline. Prebuilt ROCm archives are published for Linux and Windows, which removes the CUDA-only Python wheels from the problem entirely.
- OS
- Linux · Windows
- Stack
- Prebuilt ROCm/HIP binaries · no Python runtime
- Python runtime
- Not required
trellis.cpp — Vulkan build
The path with the fewest prerequisites: a Vulkan-capable driver replaces the entire ROCm stack. The project reports Vulkan as its fastest backend on some integrated GPUs.
- OS
- Linux · Windows
- Stack
- Vulkan driver only · no ROCm, no Python runtime
- Python runtime
- Not required
The complete cross-architecture view, including the runtime stacks and version numbers each route was tested against, is on the AMD GPU compatibility page.