Estimate — sign in for exact rates.

 Trusted GPU & AI infrastructure since 2015 — Huntsville, AL
Linux

Qwen3.8-27B UD-IQ4_XS on the Radeon RX 7900 XT 20GB (Linux)

Candidate

Candidate Linux single-card configuration for RX 7900 XT 20 GB; no published fit or performance result.

What will be tested

Card1 × AMD Radeon RX 7900 XT 20GB
Operating systemLinux
BackendNot yet selected
ModelQwen3.8-27B (Apache-2.0)
Model fileUnsloth GGUF conversion, UD-IQ4_XS
Context16,384 tokens
Concurrent requests1 (one request at a time)
PlacementFull GPU
ReviewedNot yet
Sources Upstream model · GGUF file page

Other goals for this card and OS: parallel work goal · 2 cards, two workers quality goal · 2 cards, one model split

What to expect

No published results yet. Measured results appear here only after testing and review.

Install

Candidate — no installer yet.

Overview

This Linux candidate pairs the RX 7900 XT 20 GB with the pinned model and quantization named above at the catalog's 16,384-token context. The 20 GB tier uses a compact IQ4_XS conversion of 27B at the catalog's 16K context. The point is to test a larger model choice against a smaller-quant alternative with response quality and latency measured, not to infer fit from the GGUF size. It is a test plan, not a claim of load, full-GPU placement, answer quality or speed.

Requirements

The current Linux path is a private, supervised source-checkout workflow for one recognized Vulkan GPU and a loopback server. It requires explicit artifact consent. No physical Linux AMD serving outcome is qualified. Check the card's currently reported free GPU memory, not just the number printed on its box. A longer prompt, more active requests and the desktop can change the memory budget. Keep the exact driver and operating-system build with any later test.

The pinned GGUF is 14,252,845,984 bytes (14.25 GB decimal). Keep more free local SSD space than the artifact itself for verified acquisition and cache maintenance; no fixed extra margin or host-RAM minimum has been qualified. Review the exact model license before acquisition. A planned trial must record the actual free-memory budget, selected device, loaded context, host and GPU buffers, delivered output and representative answers. The catalog context is an intended setting, not a demonstrated maximum.

Limitations

  • The exact-card load, context, GPU/host buffers, response quality and speed have no published qualification result.
  • The Linux path is a private source-checkout single-card workflow, not physically qualified on this card or released to consumers.
  • This candidate requests full-GPU placement; it is not a verified physical-residency or all-operations-on-GPU claim.

Bitworks listings for this card

Select the fields to be shown. Others will be hidden. Drag and drop to rearrange the order.
  • Image
  • SKU
  • Rating
  • Price
  • Stock
  • Availability
  • Add to cart
  • Description
  • Content
  • Weight
  • Dimensions
  • Additional information
Click outside to hide the comparison bar
Compare
Compare ×
Let's Compare! Continue shopping