Estimate — sign in for exact rates.

 Trusted GPU & AI infrastructure since 2015 — Huntsville, AL
Windows

Qwen3.8-27B Q8_0 on two Radeon RX 7900 XTX 24GB cards (Windows)

Candidate

Candidate Windows dual-split configuration for RX 7900 XTX 24 GB; no published fit or performance result.

What will be tested

Card2 × AMD Radeon RX 7900 XTX 24GB, one model split across both cards
Operating systemWindows
BackendVulkan
ModelQwen3.8-27B (Apache-2.0)
Model fileUnsloth GGUF conversion, Q8_0
Context16,384 tokens
Concurrent requests1 (one request at a time)
PlacementFull GPU
ReviewedNot yet
Sources Upstream model · GGUF file page

Other goals for this card and OS: balanced goal

What to expect

No published results yet. Measured results appear here only after testing and review.

Install

Candidate — no installer yet.

Overview

This Windows candidate proposes splitting one pinned Qwen3.8-27B Q8_0 model across two RX 7900 XTX 24 GB cards with a layer split, one active request and 16,384 context. Each card would hold a portion of the model; the result is not two independent answer workers. A split could make the Q8 artifact worth testing where one card cannot meet its estimated free-memory requirement, but transfers between cards may offset any capacity benefit. Neither aggregate advertised VRAM nor a requested full-layer count proves fit, residency or faster generation.

The 16K context comes from this exact Q8 catalog profile and its conservative reserve estimate. Some single-card Q4/Q6 profiles specify 32K, but that does not establish that this Q8 pair is limited to 16K or that the longer single-card setting would fit here. Compare matched contexts when testing latency or throughput.

Requirements

A Windows 11 lab system needs two recognized discrete cards, an AMD Vulkan-capable driver and the Microsoft Visual C++ x64 runtime. The current pair selection is heuristic, not a verified PCI/topology binding. There is no qualified dual-card launch or consumer installer. Before attempting two cards, verify the motherboard has two usable slots and record the PCIe lanes and link speed each card actually negotiates. Check the complete system's power-supply connections/capacity and case airflow against both boards; no generic PSU minimum is established here. Identify both cards individually, leave memory reserves on each, and measure sustained thermals. A pair's advertised VRAM is not one contiguous pool.

The pinned GGUF is 29,047,086,048 bytes (29.05 GB decimal). Keep more free local SSD space than the artifact itself for verified acquisition and cache maintenance; no fixed extra margin or host-RAM minimum has been qualified. Review the exact model license before acquisition. A trial must record the exact device pair and topology, each card's current free VRAM, the requested split, per-card model and compute buffers, host-memory use, API output and repeated response times. Splitting capacity and two separate replicas answer different buyer needs.

Separate Flash-Next research

Qwen3.8-Flash-Next is a different model under separate Qwen Community 1.0 terms, not this pinned 27B artifact. Every full-model conversion surveyed so far has more weight data than this pair's combined GPU memory, so Flash-Next here would depend on host-memory offload. It remains research, not a recipe. Strata's AMD HIP documentation describes Linux AMD multi-card layer splitting and host expert caching but one AMD card per model on Windows. A separate exact artifact, runtime, RAM/SSD and output-quality review is needed before any FastLLM recipe.

Limitations

  • The requested layer split and full-GPU placement have no qualified two-card load, buffer, correctness or speed result.
  • Windows pair identity is currently heuristic; PCI topology and selected-device binding need validation.
  • The Q8 16K context is the catalog candidate, not a measured maximum.
Select the fields to be shown. Others will be hidden. Drag and drop to rearrange the order.
  • Image
  • SKU
  • Rating
  • Price
  • Stock
  • Availability
  • Add to cart
  • Description
  • Content
  • Weight
  • Dimensions
  • Additional information
Click outside to hide the comparison bar
Compare
Compare ×
Let's Compare! Continue shopping