Меню

Innosilicon t2t код ошибки 35

Код Расшифровка ошибки

0 ОК

21 1 или более хеш-плат не обнаружены

22 Аномальная связь по управлению питанием

23 Все хэш-платы не могут быть включены

24 Некоторые платы не включаются

25 Не удалось поднять частоту хэш-платы

26 Не удалось установить напряжение

27 Тест чипа BIST не пройден

28 Ненормальная связь платы хешрейта не может быть автоматически восстановлена во время работы

29 Ненормальная связь по питанию во время работы не может быть восстановлена автоматически

30 Подключение к майнинговому пулу прервано

31 Повреждение отдельных микросхем, что приводит к искусственно завышенной вычислительной мощности

32 Hashboard перегрелся

33 Невозможно прочитать температуру чипа

34 Неправильное подключение кабеля связи платы управления

35 Аномальный источник питания

36 Некоторые чипы не работают должным образом

37 Тип платы управления / версия прошивки / количество микросхем не совпадает

38 Наконец, у некоторых чипов низкая вычислительная мощность.

39 Аномальные параметры старения

40 Проверьте, совпадают ли направления переднего и заднего ветра, согласуются ли они с другими машинами, и если они не совпадают, измените направление вентилятора.

41 Измерьте температуру воздухозаборника горной машины. если она превышает 40 градусов, необходимо улучшить температурную обстановку в шахте.

42 Если определенная плата вычислительной мощности часто перегревается, проблемную плату вычислительной мощности можно заменить (ремонт головоломки).

43 Ошибка, не удается прочитать температуру чипа, не удается прочитать температуру чипа, номер платы вычислительной мощности»1.Проверьте, не ослаблены ли винты на обоих концах клеммы питания и подключения кабеля SPI

44 Замените источник питания

45 Замените плату управления

46 Замените проблемную плату вычислительной мощности (ремонт головоломки)

47 Ошибка включена, неверное подключение кабеля связи платы управления, неправильное подключение кабеля SPI платы управления, номер платы вычислительной мощности «1.Проверьте, соответствует ли способ (последовательность) подключения кабеля SPI платы вычислительной мощности другим машинам той же модели

48 Замените плату управления»

49 Ошибка, неправильный источник питания, неправильный источник питания,, «1.Обратите внимание, что если нет явных отклонений в вычислительной мощности всей машины (ни одна плата не упала), нет необходимости иметь с этим дело.

50 Проверьте, не ослаблены ли винты на обоих концах клеммы питания и подключения кабеля SPI

51 Замените источник питания»

52 Ошибка, некоторые чипы работают неправильно, и количество ядер на чипе ненормально. Номер платы вычислительной мощности: Номер чипа, «1.Обратите внимание, что если нет явных отклонений в вычислительной мощности всей машины (ни одна плата не упала), нет необходимости иметь с этим дело.

53 Перезапустите майнер, чтобы узнать, по-прежнему ли сообщается о той же ошибке

54 Замените проблемную плату вычислительной мощности (ремонт головоломки)

55 ErrInvVidtype, тип платы управления/версия прошивки/количество микросхем не соответствует, тип платы управления/версия прошивки/количество микросхем не соответствует, «видтип, тип майнера, подтип, номер микросхемы», «После накопления нескольких единиц (>10) обратитесь к разработчику программного обеспечения, чтобы решить эту проблему сразу»

56 ErrBadRearChips, последние несколько чипов имеют низкую вычислительную мощность, а последние несколько чипов имеют низкую вычислительную мощность, в настоящее время не нуждаются в обработке

57 ErrInvTuneParam, параметры старения являются ненормальными, начальная частота старения и напряжение неверны, и в настоящее время старение не нуждается в обработке.

58 ,,,,,

59 ,,,,,»внимание:

60 Каждый раз, когда вы выполняете шаг решения, вам необходимо снова включить питание, чтобы убедиться, что оно вернулось в нормальное состояние.

61 После замены каждой детали, если проблема не решена, замененные детали следует вернуть на исходную машину.

62 Установлено, что отремонтированная плата вычислительной мощности требует добавления кода ошибки и простого описания проблемы».

Сегодня на Innosilicon T2TH+ 36TH/s поймал ошибку майнера 35. Как видно из журнала, контрольная плата перестала получать данные о температуре чипа 134 нулевой платы и стала выполнять попытки сброса SPI с запросом температуры 1-го чипа этой платы. Безуспешно. После 5 неудачных попыток, было принято решение о выключении (Shutdown). Напряжение с плат было снято. Прикладываю журнал сбоя «miner error code 35».

Nov 09 20:14:06 InnoMiner cgminer[997]: failed to read temperature for chain0 chip134
Nov 09 20:14:06 InnoMiner cgminer[997]: chain0: failed to read chain temperature
Nov 09 20:14:06 InnoMiner cgminer[997]: chain0: error temperature ignored(0): Tmax=9999, Tmin=9999, Tavg=9999.0
Nov 09 20:14:06 InnoMiner cgminer[997]: Chain 2 reset spihub
Nov 09 20:14:06 InnoMiner cgminer[997]: Chain 1 reset spihub
Nov 09 20:14:06 InnoMiner cgminer[997]: Chain 0 reset spihub
Nov 09 20:14:06 InnoMiner cgminer[997]: failed to read temperature for chain0 chip1
Nov 09 20:14:06 InnoMiner cgminer[997]: chain0: failed to read chain temperature
Nov 09 20:14:06 InnoMiner cgminer[997]: chain0: error temperature ignored(1): Tmax=9999, Tmin=9999, Tavg=9999.0
Nov 09 20:14:06 InnoMiner cgminer[997]: Chain 2 reset spihub
Nov 09 20:14:06 InnoMiner cgminer[997]: Chain 1 reset spihub
Nov 09 20:14:06 InnoMiner cgminer[997]: Chain 0 reset spihub
Nov 09 20:14:06 InnoMiner cgminer[997]: failed to read temperature for chain0 chip1
Nov 09 20:14:06 InnoMiner cgminer[997]: chain0: failed to read chain temperature
Nov 09 20:14:06 InnoMiner cgminer[997]: chain0: error temperature ignored(2): Tmax=9999, Tmin=9999, Tavg=9999.0
Nov 09 20:14:06 InnoMiner cgminer[997]: Chain 2 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 1 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 0 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 2 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 1 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: failed to read temperature for chain0 chip1
Nov 09 20:14:07 InnoMiner cgminer[997]: chain0: failed to read chain temperature
Nov 09 20:14:07 InnoMiner cgminer[997]: chain0: error temperature ignored(3): Tmax=9999, Tmin=9999, Tavg=9999.0
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 2 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 0 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 1 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: failed to read temperature for chain0 chip1
Nov 09 20:14:07 InnoMiner cgminer[997]: chain0: failed to read chain temperature
Nov 09 20:14:07 InnoMiner cgminer[997]: chain0: error temperature ignored(4): Tmax=9999, Tmin=9999, Tavg=9999.0
Nov 09 20:14:07 InnoMiner cgminer[997]: chain0: failed to read temperature for 5 times, SHUTDOWN
Nov 09 20:14:07 InnoMiner cgminer[997]: chain0: no power supply (0.0V)
Nov 09 20:14:07 InnoMiner cgminer[997]: miner error code: 35!
Nov 09 20:14:07 InnoMiner cgminer[997]: chain0 power down
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 2 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: Chain 1 reset spihub
Nov 09 20:14:07 InnoMiner cgminer[997]: chain2: no power supply (0.0V)
Nov 09 20:14:07 InnoMiner cgminer[997]: miner error code: 35!
Nov 09 20:14:07 InnoMiner cgminer[997]: chain2: not working due to multiple resets
Nov 09 20:14:07 InnoMiner cgminer[997]: chain2: not producing shares for more than 6 mins, SHUTDOWN
Nov 09 20:14:07 InnoMiner cgminer[997]: fan ctrl BYPASS mode enabled
Nov 09 20:14:07 InnoMiner cgminer[997]: all chains power down
Nov 09 20:14:07 InnoMiner cgminer[997]: chain1: not working due to multiple resets
Nov 09 20:14:08 InnoMiner cgminer[997]: Chain 0 reset spihub
Nov 09 20:14:09 InnoMiner cgminer[997]: failed to read temperature for chain0 chip1
Nov 09 20:14:09 InnoMiner cgminer[997]: chain0: failed to read chain temperature
Nov 09 20:14:09 InnoMiner cgminer[997]: chain0: error temperature ignored(0): Tmax=9999, Tmin=9999, Tavg=9999.0
Nov 09 20:14:09 InnoMiner cgminer[997]: Chain 0 reset spihub
Nov 09 20:14:09 InnoMiner cgminer[997]: chain0: not working due to multiple resets

После отключения произошёл нормальный запуск, питание с майнера снимать не пришлось. Откалибровался и продолжил работать без моего вмешательства.


Изменено 9 Nov 2019, 21:19 пользователем Nikolay_Po

Error code 35 innosilicon t2t

Brief Introduction

The digital coin miner is a high-end computing server and the mine farm is a kind of data center. A reliable environment with ventilation and dust resistance is required to ensure maximum hash rate, lowest failure rate and longest service life. High hash rate and returns can only be obtained in good operating environment. Therefore, never pay for your whistle just to save a small amount of money.
To maintain a good data center, you must pay attention to the temperature, humidity, dust proof and stable power. Apart from this, the miner shall be put on and taken off the shelf properly. Meanwhile, cold and heat isolation and daily inspections must be well done as well .

I. Temperature Requirement
Operation temperature: 0-40в„ѓ
However, the humidity, dust, etc. in reality will narrow the temperature range. Therefore, it is suggested to maintain the mine farm temperature between 5в„ѓ and 35в„ѓ with 25в„ѓ being the optimum temperature.

Storage temperature: -20в„ѓ-70в„ѓ

II. Humidity Requirement
Relative operation humidity: 10-90%, no condensation
Relative storage humidity:5-95%, no condensation

III. Dustproof Requirement
(1) The equipment room needs to be away from industrial pollution sources, the sea or the salt lake and free of explosive, conductive (such as metal dust lamps), magnetic and corrosive (such as sulfide, chlorine and ammonia) dusts or gases.

(2) When there is sand or dust in the equipment room, you need to install a dustproof net and clean it with a vacuum cleaner regularly.

IV. Set up the Miner Properly
1. Before setting up miners, pls check whether there are violent stubbing hints, whether the cooling fin falls off by shaking them and whether there is any appearance damage to the fan.
2. Check the wiring (for the fan and hash board) and power cable are properly connected to avoid loosening
3.The miner shall be powered off before setting them up and being removed off the shelf and handled gently. It is forbidden to be placed and dropped casually and carry the hash board wiring by hand.

V. Remove the miner off the shelf properly
1. Reconfirm whether the fault can be repaired on spot before removing the miner off the shelf.
2. Check whether the IP corresponds with the miner to avoid mismatch. You can reconfirm it by lighting the red light or reporting IP by the batch management tool.
3. If there is cold and heat isolation in data center, the air outlet where the miner is located shall be blocked to avoid hot air return in the process of removing the miner off the shelf.

4. The miner removed off the shelf shall be dustproof and moisture proof and be placed stably and orderly.

VI. Keep Power Stable:
1. The voltage of the miner socket shall be stable within the normal range of 220VВ±10%. If the voltage is too high or too low, it may cause unstable operation and burn out the power.
2. The three-phase current deviation of the power distribution cabinet shall not exceed 15%. Otherwise, the electrician shall be notified to check whether the three-phase load is balanced. The imbalance may cause voltage rise of a certain phase.
3. Check whether factory shelf and miner are grounded regularly (It is requested that the grounding resistance be lower than 4О©.). If they are not grounded or well grounded, pls notify the professional electrician to finish the grounding work in time. If you often feel tingling by touching the miner in the process of operation and maintenance, pls check whether the grounding work is done or done properly.
4. The cable cannot be placed in the hot air area. Pls check regularly the cable aging condition.
5. Try to avoid frequent power outage in the miner farm. Planed outage is necessary. If you need to cut the power off, pls turn off the automatic air switches on the shelf by power in ascending order and finally the master switch. Before powering on, pls make sure all the air switches connecting the miner on the shelf are turned off. The air switches shall be turned on in descending order from the master switch to avoid any possible damage caused by transient voltage surge.

VII. Good Job of Cold and Hot Isolation
Cold and hot isolation must be done in the mine farm especially in the ones with high power miners.

VII. Good Job of Cold and Hot Isolation
Precautions for cold and hot isolation:

VIII. Daily Inspections in Operation and Maintenance:
You shall check the miner and mine farm in the respect of temperature, humidity, miner appearance, the environment condition and power condition. Pls refer to the attached table for detailed inspection items.

VIII. Daily Inspections in Operation and Maintenance:

Good execution of the listed contents above at the same time and offering your high-end server a good operating environment are powerful measures to ensure your miner high hash rate and high returns!

Q1пјљHow to locate the faulty part by the error code of the miner?

(1) Download the batch management tool to the desk and open it. The downloading link is as follows:
http://www.innosilicon.com.cn/download/InnoMonitor%20v1.0.5_beta.rar
(2) Input the IP range to which the miner belongs and click the “Scan” icon.

If there is indeed a problem with the miner but no error code is reported by the batch management tool, go to the backstage log screen of the miner to look for the error code:

(3)Check the miner fault according to the error code and confirm the fault location.
Notes:
1. The miner shall be powered on again to see whether it becomes normal each time you perform one resolution step.
2. The replaced part shall be reinstalled on the original miner once again after replacing the part but with the default unsettled.
3. The miner needs repairing shall be pasted with the error code and brief problem description.

Источник

Innosilicon A11

mineun

Новичок

Здравствуйте. Уже вторую неделю борюсь с проблемой : пропала хэш плата с 24 ошибкой

КОД ОШИБКИ:
24
сообщение об ошибке:
ENCORE_FAIL.цепь(4).чип().msg()
описание:
Некоторые платы питания не могут быть включены
решение:
1.Проверьте, не ослаблены ли винты на обоих концах клеммы питания и подключения кабеля SPI
2.Замените источник питания
3.Замените плату управления
4.Замените проблемную плату вычислительной мощности (ремонт головоломки)
внимание:
1.Каждый раз, когда вы выполняете шаг решения, вам необходимо снова включить питание, чтобы подтвердить, вернулось ли оно в нормальное состояние.
2.После замены каждой детали, если проблема не решена, замененные детали следует вернуть на исходную машину.
3.Определено, что переработанная плата вычислительной мощности требует прикрепления кода ошибки и простого описания проблемы.

Посмотрел видео, где такая же проблема — вернул заводские настройки, перезагрузил, дождался автотюна, — по прежнему 3 платы видит из 4, потом почистил клеммы, оставил на ночь
Утром включаю — заработали 3 платы, на следующий день — выключили поставили в шумобокс.
Когда включили опять тоже самое.

куда копать? платы местами менять?

Senechkin a10pro

Пляшущий с бубном

Здравствуйте. Уже вторую неделю борюсь с проблемой : пропала хэш плата с 24 ошибкой

КОД ОШИБКИ:
24
сообщение об ошибке:
ENCORE_FAIL.цепь(4).чип().msg()
описание:
Некоторые платы питания не могут быть включены
решение:
1.Проверьте, не ослаблены ли винты на обоих концах клеммы питания и подключения кабеля SPI
2.Замените источник питания
3.Замените плату управления
4.Замените проблемную плату вычислительной мощности (ремонт головоломки)
внимание:
1.Каждый раз, когда вы выполняете шаг решения, вам необходимо снова включить питание, чтобы подтвердить, вернулось ли оно в нормальное состояние.
2.После замены каждой детали, если проблема не решена, замененные детали следует вернуть на исходную машину.
3.Определено, что переработанная плата вычислительной мощности требует прикрепления кода ошибки и простого описания проблемы.

Посмотрел видео, где такая же проблема — вернул заводские настройки, перезагрузил, дождался автотюна, — по прежнему 3 платы видит из 4, потом почистил клеммы, оставил на ночь
Утром включаю — заработали 3 платы, на следующий день — выключили поставили в шумобокс.
Когда включили опять тоже самое.

куда копать? платы местами менять?

Я бы делал так.
1) Проверил блок питания.
2) Проверил бы приходящие ток на плату которую не видит.
3) Поменял платы местами если плата рабочая. (то вероятно проблемы с контрольной платой)

Ну короче по логичной цепочке пошёл бы.
Приходит не приходит ток, работает не работает плата.
Тока нет, плата работает на другом месте. Логично идём по цепи к контрольной плате.

Источник

Руководство по ремонту хэш-платы майнеров Innosilicon [EN]

Document Type: Maintenance Plan

Contents of this booklet: Mainly describes how to troubleshoot various faults of the T1.T2 hash board and how to use the test to accurately locate.

Scope: applicable to all T1 production, after‑sales, and outsourcing maintenance sites

1. Maintenance platform requirements:

1. Constant temperature soldering iron ( 350 Degree‑‑ 400 Degrees), the pointed soldering iron tip is used for soldering small patches such as chip resistors and capacitors. Skilled mastery.

2. The hot air cylinder is used for chip disassembly and soldering. Be careful not to heat it for a long time to avoid PCB foaming. 

3. DC stabilized power supply (output 12V, 20A), used for the test and measurement of the hash board. 

4. Fluke 15b+ multimeter, tweezers, Debug, G7 maintenance special control board, oscilloscope.

5. Flux solder paste, washing water and absolute alcohol; washing water is used to clean up the solder residue and appearance after repair.

6. Tin planting fixture, planting tin steel mesh, solder paste; when replacing a new chip, you must plant the chip with tin.

7. The thermal conductive glue is black, high temperature, gray low temperature used for re-attaching the heat sink after maintenance.

2. Requirements on Maintenance Operations: 

1. The maintenance personnel must have certain electronic knowledge, more than one year of maintenance experience, and master QFN package welding technology.

2. After repairing, the hash board must be tested twice and confirmed as OK before it can pass! 

3. Pay attention to the operation method when replacing the chip. After replacing any accessories, the PCB board is not obviously deformed, and the replaced parts and the surrounding area shall be checked for whether there is open and short circuit.

4. Determine the maintenance station object and the corresponding test software parameters and test fixtures. 

5. Check whether the tools and jigs can work normally.

3. Principle and structure:

● Principle overview

1. T1 is composed of 21 voltage domains in series, each voltage domain has 3 chips, and the whole board has 63 T1558 chips.

2. The T1558 clock is two 12M crystal oscillators, which are transmitted in series from the first chip to the 30th, and 31 to the last chip.

3. There is an independent small heat sink on the back of each chip of T1. The small heat sink on the back is fixed on the back of the IC with thermal glue after the initial test of the board. Repair and replace the chip after passing the test, you need to evenly apply black thermal conductive glue on the IC surface and heat it to fix it.

● Analysis of key points:

The following figure shows the SPI trend and voltage domain of the PCB board and the chip sequence bit number.

repair hashboard

figure 1

Test whether the SPI waveform of the error‑reporting chip is normal.

1. Each yellow box in the figure is a voltage domain, a total of twenty one Voltage domains, each voltage domain is on average 0.42V.

2. The black numbers represent the order and bit number of the chip.

3. The red arrow in the figure shows CLK Signal direction.

The yellow arrow shows the direction of the SCK signal;

The green arrow shows the direction of the CS signal;

The blue arrow shows the direction of the DI signal;

The purple arrow shows the direction of the DO signal.

4. There is between every two chips 1‑7 Test point 1 for CLK Signal; test point 2 for RST Signal; test point 3 for EN Signal; test point 4 for SCK Signal; test point 5 for CS Signal; test point 6 for DI Signal; test point 7 for DO signal.

DI signal flow direction, from No. 63 chip to 1 Return the chip number, and then return to the control board;

DO signal flow direction, by 1 No. chip pulls low level toward 63; not plugged in IO Line, standby 0V , When calculating 0.3 Pulse signal around.

The RST signal flows in from the control board, and then by 1 Chip to 63 No. chip transmission.

2.2 The figure below shows the key circuits on the front of the T1 hash board.

repair hashboard

figure 2

1). Test points between each chip (as shown in the figure after zooming in): Figure 2

Figure 2. When repairing test points between chips, the test points between test chips are the most direct way to locate faults. The arrangement of the test points of the T1 arithmetic board is: CLK, RST.EN, SCK, CS, DI, DO signals.

Figure 1. Signal trend

2) Voltage domain: The whole board has 21 voltage domains, and each voltage domain has 3 chips. The three chips in the same voltage domain are powered in parallel, and then connected in series with other voltage domains after being connected in parallel. The circuit structure is shown in Figure 4 below:

Principle analysis of voltage domain single chip (see Figure 3 below)

maintenance Innosilicon hashboard

Figure 3

● The above are the functions of each pin of the T1558 chip.

During maintenance, 14 test points before and after the chip are mainly tested (seven points before and after the chip: CLK, RST, EN, SCK, CS, DI, DO); DCDC voltage output 8.82V; boost voltage 11V, LDO—1.8 V etc.

maintenance Innosilicon hashboard

The two ends of the C56 capacitor on the left are the total DCDC output voltage, which should be about 8.82V

maintenance Innosilicon hashboard

On the left, both ends of the C57 capacitor are the boost voltage, which should be about 11V

BM1558 circuit diagram

Figure 6. BM1558 circuit diagram

BM1760 chip pins

Figure 7. BM1760 chip pins

CLK: 0.9V provided by Y1 12M crystal oscillator;

DO: From the first chip to the last chip provided by the control board, the signal can be measured with an oscilloscope;

DI: Return from the last chip to the first chip, the signal can be measured with an oscilloscope;

SCK: When the control board provides about 0.12V for calculation, the abnormal or low voltage will cause the calculation board to be abnormal or the calculation power is low;

EN: 1.8V Provided by the control board;

CS: Provided by the control board;

RST: 1.8V . Provided by the control board, each time the test key is pressed, a low‑level reset signal will be output again.

When the above‑mentioned test point status and voltage are abnormal, please estimate the fault point based on the circuit before and after the test point.

It can be seen from the chart above:

CLK signal: by the chip 32 Or 31‑pin in, 17‑pin out, when connected across the voltage domain, by 5 Foot out through 100NF The capacitor is connected to the input to the next chip twenty three foot.

DO signal: enter from pin 6 of the chip, 12 Foot out;

DI signal: returned by the chip from pin 5, output from pin 13 or 14;

CS signal: input from pin 3 of the chip and output from pin 15;

RST signal: Input from chip 30 pins, output from 118 pins.

Test the signal voltage of each chip, LDO‑1.8OV

CORE: 0.8V When this voltage is abnormal, it is usually the chip of the voltage domain CORE Short circuit

LDO‑1.8O: 1.8V When this voltage is abnormal, the chip LDO‑1.8O Short circuit or open circuit

3) Judging the operating status of the hash board, the hash rate of the chip, and the temperature sensitivity based on the information in the printing window of the manufacturing tool.

3.3 IO Interface Definition

IO is composed of 2X7 pitch 2.0 PHSD 90 degree in‑line double row.

Innosilicon hashboard repair manual

The pin definitions are shown in Figure 8 below:

As shown in FIG:

1 pin is LED

2 pin for VIDD

10, 14 pin: for GND .

3 pin are EN

4 pin for STAR

7 pin is PLUG

12 pin for SCK

13 pin are CS

8.9 pin (DI, EO)

6 pin ( RST ): is the reset signal 3.3V Terminal, after being divided by resistors, it becomes 1.8V RST Reset signal.

5 pin ( 3V3 ): is the hash board 3.3V Power supply, the 3.3V Provided by the control board, mainly for PIC Provide working voltage.

Figure 8. IO Definition of each pin

TX_IN voltage is 1.8V

RST_IN voltage is 1.8V

4. Routine maintenance process:

● Reference steps:

1. Routine inspection: First, perform visual inspection on the arithmetic board to be repaired to observe whether there is any displacement, deformation, or scorching of the small heat sink? If any, you must deal with it first; if the small heat sink is displaced, remove it first, wash off the original glue, and re‑adhesive after the repair is passed.

Secondly, after the visual inspection is no problem, the impedance of each voltage domain can be tested first to detect whether there is a short circuit or an open circuit. If you find out, you must deal with it first.

Thirdly, check whether the voltages in each voltage domain reach 0.4v, and the voltage difference between the voltage domains must not exceed 0.05. If the voltage in a voltage domain is too high or too low, the circuits in the adjacent voltage domain generally have abnormal phenomena, and it needs to find the reason first.

2. After the routine detection is no problem (generally, the short‑circuit detection of the routine detection is necessary, so as not to burn the chip or other materials due to the short circuit when the power is turned on), you can Use DEBUG connection for chip detection, and judge and locate according to the detection result.

3. According to the display results of the test and detection, starting from the vicinity of the faulty chip, check the chip test points (CLK, RST, EN, SCK, CS, DI, DO); DCDC voltage output 8.82V; boost voltage 11V, LDO‑‑ 1.8V etc.

4. Then according to the signal flow direction, except for DI signals, the signals are transmitted in the reverse direction (chips 6 to 1). Several of the signals CLK, RST, EN, SCK, CS, DO are forward transmission (1-63), and abnormalities are found through the power supply sequence the point of failure.

5. When locating the faulty chip, the chip needs to be welded again. The method is to add flux around the chip (preferably no-clean flux), heat the solder joints of the chip pins to a dissolved state, move gently up and down, left and right to press the chip; prompting the chip pins and pads Re-melt and collect tin. In order to achieve the effect of tinning again. If the fault remains the same after re-soldering, you can directly replace the chip.

6. The repaired arithmetic board must be tested twice or more during testing. Two test times before and after: for the first time, after the replacement of parts is completed, the hash board needs to be cooled down, and after passing the test, put it aside first. For the second time, after a few minutes wait for the arithmetic board to cool down completely, perform the test again. Although the time for the two tests is a few minutes, this does not affect the work. Put the repaired board aside, continue to repair the second board, wait for the second board to be repaired and set it aside to cool down, and then test the first board. In this way, the time is just staggered, and the total time is not delayed.

7. The repaired board. It is necessary to classify the faults and make records of the type, location, reason, etc. of the replacement components. For feedback back to production and after‑sales, Research and development.

8. After recording, install it into a complete miner for formal aging.

5. Failure types:

1. The impedance of each voltage domain is unbalanced; when the impedance of certain voltage domains deviates from the normal value, it indicates that there are parts in the abnormal voltage domain that have open circuits and short circuits. It is most likely to be caused by general chips. But there are three chips in each voltage domain, and often only one has a problem when it fails. The method of finding out the problem chip can detect and compare the abnormal point through the test point to ground impedance of each chip. If you encounter a short-circuit phenomenon, you can first remove the heat sink on the chip with the same voltage, and then observe whether the chip pins are connected to the solder. If the short-circuit point cannot be found in the appearance, the short-circuit point can be found according to the resistance method or the current interception method.

2. Voltage imbalance in the voltage domain; When the voltage of some voltage domains is too high or too low, it is generally because of abnormal voltage domains or adjacent voltage domains that there are abnormal signals, resulting in abnormal working status of the next or next voltage domain and voltage imbalance . The abnormal point can be found only by detecting the signal and voltage of each test point. Individually, it is necessary to find out the abnormal point by comparing the impedance of each test point.

Observe the appearance, measure the impedance, measure the voltage, and check the voltage and power supply of each test point. The test locates the chip according to the test information, first re-soldering, and re-soldering is invalid. The fault type is recorded and tested for more than two times. Ok can be considered as repaired, and then related aging.

Pay special attention to the fact that the CLK signal and RST Signal, these two abnormalities are most likely to cause voltage imbalance.

3. Lack of chips: The lack of chips means that the test box fails to detect all 63 chips, often only as many as the actual number of chips. However, the actual missing (undetected) abnormal chip is not in the displayed position. At this time, it is necessary to accurately locate the abnormal chip through testing. The location method can use TX cut-off to send out the way to find the location of the abnormal chip. It is to connect the TX signal of a certain chip to the ground. For example, after outputting the TX signal of the 50th chip to the ground of the voltage domain, theoretically if all the previous chips are normal, 50 chips should be detected in the test box? If 50 chips are not detected, the abnormality is before the 50th chip; if 50 chips are detected, the abnormality is after the 50th chip. By analogy, use the dichotomy to find the location of the abnormal chip.

4. Broken chain:

A broken chain is similar to lack of chips, but in a broken chain, not all chips that cannot be found are abnormal, but all the chips after the abnormal chip are invalid due to a certain chip abnormality. For example, a chip itself can work, but it will not forward other chip information; at this time, the entire signal chain will come to an abrupt end, and lose a large part of it, which is called broken chain. Generally the broken chain can be displayed by the test box. For example, when the test box detects the chips, only 14 chips are detected. If the number of preset chips is not detected in the test box, it will not run, so it will only display how many chips are detected, at this time, according to the displayed number «14», the problem can be found by detecting the voltage and impedance of each test point before and after the 14th chip.

5. Not running:

No running means that the test box cannot detect the chip information of the hash board, but displays NO hash board; this phenomenon is the most common and the fault range involved is also wide.

1) Non‑operation caused by abnormal voltage in a certain voltage domain; the problem can be found by measuring the voltage of each voltage domain.

2) The abnormality caused by a certain chip abnormality can be found by measuring the signal of each test point.

CLK signal: the signal is generated by 1 No. chip output to 63 No. chip, but the current version has only two crystal oscillators, Y1(1‑30) X1(31‑63) of which as long as there is an abnormal signal clk  Yes, all the following signals will be abnormal, search in order according to the signal transmission direction.

DO signal: This signal is caused by 1 , 2 , 3 ,,,,, 63 No. chip, when a certain point of the dichotomy is abnormal, it can be detected forward.

DI signal: This signal is returned by No. 63.60, 59, 58, and 1, and the cause of the fault is confirmed through the chip signal direction. This signal is the highest priority if the T1 operation board is not running, and the signal is searched first.

RST signal: 1.8V ; After the arithmetic board is powered on and the 14P signal is plugged in, this signal will change from 01 , 02 ,,,,,, 0 63 The direction of the transmission to the last chip.

3) A certain chip VDD It can be caused by measuring whether the potential difference of each voltage domain is normal. Under normal circumstances, when the VDD voltage is 0.42, the normal voltage of each test point in other voltage domains is also 0.42 to ensure the balance between the voltage domains.

4) of a certain chip VDD1V8 Abnormal voltages Determine whether a certain VDD1V8 voltage is normal by measuring the test points of each voltage. Generally, the LDO voltage determines the voltage of each test point. When the LDO voltage is 1.8V, the normal voltage of each test point in other voltage domains is also 1.8. V

5. Low hash rate:

Low hash rate can be divided into:

1) During the test, received Nonce Insufficient, lack of hash rate and show bad phenomena. This phenomenon can be judged by seeing the number of nonce returned by each chip directly through the serial port printing information. Generally, the chip with the returned nonce number lower than the set value should be trouble‑shooted, and the non‑welding and external causes can be directly replaced. .

2) When the test fixture was tested, the hash rate was low after the whole miner was installed. Most of this situation is related to the heat dissipation conditions of the chip, and special attention should be paid to the glue used for the small heat sink of each chip and the ventilation performance of the whole miner. Another reason is that the voltage of a certain chip is critical. After the whole miner is installed, the difference between the 12V power supply and the power supply during the test causes the test calculation power to deviate from the running calculation power. You can use the test box to test after turning it down, and adjust it slightly. After the 12V output of the voltage DC adjustable power supply, perform the test again and find out the voltage domain with the lowest number of returned nonces.

6. A certain chip NG:

Refers to when the test is passed, the test serial port information shows that the returned nonce of a certain chip is insufficient or zero. In addition to eliminating the problem of false soldering and peripheral components, you can Replace the chip directly.

● Maintenance  instructions:

1. During maintenance, the maintenance personnel must be familiar with the function and flow direction of each test point, the normal voltage value and the ground impedance value.

2. You must be familiar with chip soldering to avoid blistering and deformation of the PCB or damage to the pins.

3. T1558 chip package, 16 pins on both sides of the chip. The polarity and coordinates must be aligned during welding and they must not be misaligned.

4. When replacing the chip, the thermally conductive fixing glue around the chip must be cleaned to prevent the chip from being damaged by the hanging or poor heat dissipation when the IC is soldered.

newbie

Activity: 39

Merit: 0

Hi,

I can’t seem to access my miner logs…. how can i do it?

legendary

Activity: 2744

Merit: 2373

Merry Up Christmas and Happy New Year to All.

I think it would be better if you can share with us your mlogs here so that we can find what error do you get or what is the reason before the error code popup.

Look at this image as a sample:

source: https://www.innosilicon.com/html/support_en/q_a.html

As you can see there is an error code 35 and according to innosilicon error code it’s no power output and look at the image before error code 35 «chain1: no power supply -0.0v».

So check the mlogs and post it here in full so that we can see the issue before the error 161.

legendary

Activity: 1750

Merit: 4808

be constructive or S.T.F.U

Are you sure it’s 161? as far as I know, the error codes don’t go past 40 something, can you post a screenshot? also have you tried to flash a different firmware on one of the gears that shows this error?

newbie

Activity: 39

Merit: 0

Hi,

Anyone have any idea what is error code 161? I tried googling but can’t seem to find what is the issue.

It shows error code 161 on quite a few rigs but it is still mining.

0 0 голоса
Рейтинг статьи
Подписаться
Уведомить о
guest

0 комментариев
Старые
Новые Популярные
Межтекстовые Отзывы
Посмотреть все комментарии

А вот еще интересные материалы:

  • Яшка сломя голову остановился исправьте ошибки
  • Ясность цели позволяет целеустремленно добиваться намеченного исправьте ошибки
  • Ясность цели позволяет целеустремленно добиваться намеченного где ошибка
  • Injusticelauncher exe системная ошибка
  • Injustice ошибка при запуске приложения 0xc0000906