---
title: AMD Radeon RX 9060 XT 8GB
url: "https://bitworks.io/amd-inference/rx-9060-xt-8gb/"
description: "The 8 GB RX 9060 XT has half the advertised memory of its 16 GB sibling. Its candidate Qwen3.5-9B Q4_K_M conversion uses a shorter catalog context to reduce the initial memory demand. That is a planning choice, not a demonstrated maximum context or proof of usable speed. Check the memory reported free when inference starts, […]"
---

# AMD Radeon RX 9060 XT 8GB

> The 8 GB RX 9060 XT has half the advertised memory of its 16 GB sibling. Its candidate Qwen3.5-9B Q4_K_M conversion uses a shorter catalog context to reduce the initial memory demand. That is a planning choice, not a demonstrated maximum context or proof of usable speed. Check the memory reported free when inference starts, […]

Human-readable page: https://bitworks.io/amd-inference/rx-9060-xt-8gb/

## Content

The 8 GB RX 9060 XT has half the advertised memory of its 16 GB sibling. Its candidate Qwen3.5-9B Q4_K_M conversion uses a shorter catalog context to reduce the initial memory demand. That is a planning choice, not a demonstrated maximum context or proof of usable speed.

Check the memory reported free when inference starts, particularly if the card also drives a display. The model file, KV state and compute buffers all need space. If a task needs better answer quality or a longer conversation, compare the 16 GB variant's distinct Q8 recipe only after testing both exact artifacts. This 8 GB page has no published physical result.

For agents: how to buy here — https://bitworks.io/checkout.md
