# Five used mining boards run Qwen3-Coder-Next at 40 tokens per second

> A Reddit user says a cluster of five BC-250 boards generates Qwen3-Coder-Next at 40 tokens per second. The setup reportedly cost less than $800, but uses power inefficiently.

Oossa · 2026-09-29 · https://oossa.com/en/five-used-mining-boards-run-qwen3-coder-next-at-40-tokens-per-second

A LocalLLaMA Reddit user has put five used BC-250 mining boards together to run Qwen3-Coder-Next, an AI model for writing code. They report about 40 tokens per second with a 30,000-token context, slowing to around 30 tokens per second at 100,000 tokens.

The boards share about 71GB of video memory and communicate over their built-in 1-gigabit Ethernet. The builder says the setup cost less than $800, while noting that it is “wildly inefficient” with power; they may add two more boards to test another model.

## The facts

- The five-board cluster reportedly reaches 40 tokens per second at a 30,000-token context.
- The builder says the setup cost less than $800 and exposes about 71GB of video memory.

## Why it matters

The build shows how used mining hardware can be repurposed for local AI, though its power use may limit its practicality.

## Sources & references

1. [I’m calling this the Monstrosity. 5 ex mining BC-250 boards Qwen3-Coder-Next Q4 at 40 tok/s](https://www.reddit.com/r/LocalLLaMA/comments/1ws49si/im_calling_this_the_monstrosity_5_ex_mining_bc250/) – Reddit r/LocalLLaMA, 2026-09-28

Last updated: 2026-09-29
