Categories
Mining Education, Mining Guides

Learn how to diagnose and fix the most common ASIC miner hardware failures including hashboard errors, fan malfunctions, PSU issues, and network connectivity problems to minimize downtime and maximize hashrate.

When an ASIC miner goes down, every hour of downtime is lost revenue. Whether you are hosting miners at a colocation facility or running hardware at home, knowing how to diagnose and resolve common hardware failures can save you thousands of dollars in repair costs and lost hashrate. This guide walks through the most frequent ASIC miner problems, their root causes, and practical steps to get your machines hashing again.

Hashboard Failures: The Most Common ASIC Problem

Hashboard failures account for roughly 30 to 40 percent of all ASIC miner downtime. A single failed hashboard on a three-board machine like the Antminer S21 means you lose approximately one-third of your hashrate while still drawing significant standby power.

Symptoms of a Failing Hashboard

  • Reduced hashrate — Your miner reports significantly lower hashrate than its rated output. For example, an S21 Pro running at 140 TH/s instead of 234 TH/s suggests one board is offline.
  • Missing board in dashboard — The miner’s web interface or firmware dashboard shows only two of three boards detected.
  • Elevated chip temperatures on remaining boards — Remaining boards may run hotter as the control system compensates.
  • Frequent restarts — The miner repeatedly reboots as it tries to initialize a defective board.

Diagnosing the Root Cause

Start with the miner’s kernel log. Access it through the web interface under System > Kernel Log or via SSH. Look for error messages referencing specific chain numbers (chain 0, chain 1, chain 2). Common error patterns include:

  • “Chain [X] only found 0 chips” — The board is not communicating. This often indicates a failed ASIC chip, broken data line, or loose ribbon cable connection.
  • “Temp sensor lost” — Temperature monitoring has failed on a board, which triggers a safety shutdown.
  • “Voltage error on chain [X]” — Power delivery to the hashboard has failed, possibly due to a blown voltage regulator or damaged power connector.

What You Can Fix Yourself

  1. Reseat the ribbon cables and power connectors. Power off the miner completely, open the enclosure, and firmly reseat every cable. Loose connections cause a surprising number of hashboard “failures.”
  2. Inspect for physical damage. Look for burn marks, swollen capacitors, or corroded solder joints on the hashboard. Visible damage usually means the board needs professional repair.
  3. Swap boards between miners. If you have two identical miners, swap a suspect board into the working machine. If the problem follows the board, you have confirmed the board is defective.

Fan Errors and Cooling Failures

ASIC miners run extremely hot, and fan failures can cause thermal shutdowns within minutes. Modern miners like the S21 series use four high-speed fans rated at 6,000 to 7,000 RPM. When a fan fails or slows down, the miner’s firmware triggers a protective shutdown to prevent chip damage.

Common Fan Error Scenarios

  • Fan speed at 0 RPM — The fan has completely failed. Check the power connector first. If the connector is secure, the fan motor has likely burned out and needs replacement.
  • Fan speed below minimum threshold — Dust buildup on fan blades or bearing wear can reduce RPM below the firmware’s minimum acceptable speed. Clean the fan with compressed air and test again.
  • All fans reading abnormal speeds — If every fan shows the same unusual reading, the control board (not the fans) may be at fault. A firmware reflash sometimes resolves this.

Preventive Measures

Replace fans proactively every 12 to 18 months in high-dust environments. Stock spare fans for your most common miner models. At a hosting facility, confirm that your provider monitors fan health and has replacement parts on hand.

Power Supply Unit (PSU) Issues

The PSU converts AC power from the wall into the DC voltage your hashboards need. PSU failures can manifest in several ways:

  • Miner will not power on — No lights, no fans, no response. Check the AC input cable and outlet first. If power is confirmed at the outlet, the PSU’s internal fuse or main switching circuit may have failed.
  • Intermittent shutdowns — The miner runs for minutes or hours, then shuts down unexpectedly. This often indicates an overheating PSU (check for blocked ventilation) or a PSU operating at the edge of its rated capacity.
  • Voltage out of range — If the miner’s interface reports input voltage outside its acceptable range, check your facility’s electrical supply. Voltage sag during peak demand periods is a common issue at sites with undersized transformers.

PSU Safety Warning

Never open a PSU enclosure. PSU capacitors store lethal voltages even when unplugged. If a PSU has failed, replace it entirely. For hosted miners, contact your colocation provider to handle PSU replacements.

Network and Connectivity Problems

Miners need stable internet connectivity to communicate with their mining pool. Network issues do not damage hardware but they destroy your effective hashrate.

  • High reject rate — If your pool dashboard shows a reject rate above 2 percent, check your network latency to the pool server. Use a different pool server in a closer geographic region if latency exceeds 100 milliseconds.
  • Frequent pool disconnections — Replace the Ethernet cable first. Cat5e cables in hot, dusty environments degrade faster than expected. Ensure your router or switch port is functioning correctly.
  • DHCP conflicts — Large mining farms with hundreds of devices can exhaust DHCP address pools. Assign static IP addresses to miners to eliminate this issue.

Firmware and Software Glitches

Sometimes the problem is not hardware at all. Corrupted firmware or configuration errors can mimic hardware failures.

  • Miner stuck in a boot loop — A corrupted firmware image can prevent the miner from completing startup. Most manufacturers provide SD card recovery images that reflash the control board firmware.
  • Hashrate significantly below spec after firmware update — Roll back to the previous firmware version. Not all firmware updates improve performance on every hardware revision.
  • Custom firmware causing instability — If you are running Braiins OS+, LuxOS, or Vnish, ensure the firmware version matches your specific hardware model and revision. Mismatched firmware is a frequent source of instability.

When to Call a Professional Repair Service

Some repairs require specialized equipment including hot air rework stations, oscilloscopes, and replacement ASIC chips. Send your hashboard to a professional repair service when:

  • You see visible burn marks or damaged components on the PCB
  • Multiple ASIC chips on a single board are reporting errors
  • The board passes visual inspection but still fails to initialize
  • You lack the soldering equipment to perform BGA rework

Professional hashboard repair typically costs between 100 and 300 dollars per board, far less than replacing the entire miner. For hosted miners, ask your provider if they offer on-site repair services or have relationships with repair shops.

Building a Troubleshooting Toolkit

Every miner operator should keep these items on hand:

  • Spare Ethernet cables (Cat6 rated)
  • Compressed air cans or an electric duster
  • Replacement fans for your most common miner models
  • A multimeter for checking AC voltage and PSU output
  • Thermal paste (for reapplication during maintenance)
  • SD card with recovery firmware images
  • Anti-static wrist strap

Frequently Asked Questions

How do I know if my hashboard or PSU is the problem?

Swap the suspected hashboard into a working miner with a known-good PSU. If the board works in the other machine, your original PSU is likely the issue. If the board fails in both machines, the hashboard itself is defective.

Can I run my ASIC miner with a failed hashboard?

Yes, most miners will operate with one or two working hashboards, though at reduced hashrate. However, running long-term with a failed board is not recommended because the remaining boards may compensate by drawing more power, stressing the PSU.

How often should I clean my ASIC miners?

In a clean data center environment, every three to six months is sufficient. In dusty or outdoor environments, monthly cleaning may be necessary. Excessive dust is the leading cause of fan failures and overheating.

Should I attempt repairs myself or use a hosting provider with maintenance services?

For cable reseating, fan replacements, and firmware reflashes, self-service is straightforward. For component-level board repairs, use a professional service. If you host with a colocation provider, ask about their maintenance and repair policies before signing your agreement.

What is the most common cause of ASIC miner failure?

Heat and power quality are the two leading causes. Insufficient cooling leads to thermal stress that degrades ASIC chips over time, while voltage spikes or sags from unstable power supplies damage sensitive electronic components. Hosting at a professional facility with proper electrical infrastructure and industrial cooling significantly reduces these risks.

Keeping your miners running at peak performance starts with knowing what to look for when something goes wrong. Combine regular preventive maintenance with a systematic troubleshooting approach, and you will minimize downtime and maximize your mining revenue. Contact our team if you need help diagnosing or resolving hardware issues with your hosted miners.

Explore Rax Mining

Categories