메뉴 건너뛰기

Cloudera, BigData, Semantic IoT, Hadoop, NoSQL

Cloudera CDH/CDP 및 Hadoop EcoSystem, Semantic IoT등의 개발/운영 기술을 정리합니다. gooper@gooper.com로 문의 주세요.


bin/start-hbase.sh을 실행하면 아래와 같은 오류가 발생하면서 HMaster 데몬이 기동하지 않는 경우있어,

제시되는 "hbase hbck -fixVersionFile"를 실행하면 실행로그처럼 루프에 빠지게 되는 경우가 있다.

이는 데이타 유실이 발생되어 복구할 수 없는 상태를 나타내므로 zookeeper와 HDFS의 hbase관련 정보및 데이타를 모두 삭제하여야 한다.

(./hbase hbck -fixMeta -fixAssignments를 실행할때도 loop에 빠지게 되면 zookeeper의 /hbase노드를 삭제하고 HDFS상의 /hbase를 삭제한후 bin/start-hbase.sh를 실행한다.)


----hbase hbck -fixVersionFile실행로그---

2016-08-01 14:55:02,619 INFO  [Group Metadata Manager on Broker 1]: Removed 0 expired offsets in 1 milliseconds. (kafka.coordinator.GroupMetadataManager)
2016-08-01 14:55:16,781 INFO  [main] client.RpcRetryingCaller: Call exception, tries=10, retries=35, started=48182 ms ago, cancelled=false, msg=
2016-08-01 14:55:36,832 INFO  [main] client.RpcRetryingCaller: Call exception, tries=11, retries=35, started=68233 ms ago, cancelled=false, msg=
2016-08-01 14:55:56,942 INFO  [main] client.RpcRetryingCaller: Call exception, tries=12, retries=35, started=88343 ms ago, cancelled=false, msg=
2016-08-01 14:56:17,080 INFO  [main] client.RpcRetryingCaller: Call exception, tries=13, retries=35, started=108481 ms ago, cancelled=false, msg=
2016-08-01 14:56:37,203 INFO  [main] client.RpcRetryingCaller: Call exception, tries=14, retries=35, started=128604 ms ago, cancelled=false, msg=
2016-08-01 14:56:57,328 INFO  [main] client.RpcRetryingCaller: Call exception, tries=15, retries=35, started=148729 ms ago, cancelled=false, msg=
2016-08-01 14:57:17,510 INFO  [main] client.RpcRetryingCaller: Call exception, tries=16, retries=35, started=168911 ms ago, cancelled=false, msg=
2016-08-01 14:57:37,631 INFO  [main] client.RpcRetryingCaller: Call exception, tries=17, retries=35, started=189032 ms ago, cancelled=false, msg=
2016-08-01 14:57:57,818 INFO  [main] client.RpcRetryingCaller: Call exception, tries=18, retries=35, started=209219 ms ago, cancelled=false, msg=
2016-08-01 14:58:17,979 INFO  [main] client.RpcRetryingCaller: Call exception, tries=19, retries=35, started=229380 ms ago, cancelled=false, msg=
2016-08-01 14:58:38,165 INFO  [main] client.RpcRetryingCaller: Call exception, tries=20, retries=35, started=249566 ms ago, cancelled=false, msg=
2016-08-01 14:58:58,282 INFO  [main] client.RpcRetryingCaller: Call exception, tries=21, retries=35, started=269683 ms ago, cancelled=false, msg=
2016-08-01 14:59:18,410 INFO  [main] client.RpcRetryingCaller: Call exception, tries=22, retries=35, started=289811 ms ago, cancelled=false, msg=
2016-08-01 14:59:38,572 INFO  [main] client.RpcRetryingCaller: Call exception, tries=23, retries=35, started=309973 ms ago, cancelled=false, msg=
2016-08-01 14:59:58,699 INFO  [main] client.RpcRetryingCaller: Call exception, tries=24, retries=35, started=330100 ms ago, cancelled=false, msg=
2016-08-01 15:00:18,800 INFO  [main] client.RpcRetryingCaller: Call exception, tries=25, retries=35, started=350201 ms ago, cancelled=false, msg=
2016-08-01 15:00:38,807 INFO  [main] client.RpcRetryingCaller: Call exception, tries=26, retries=35, started=370208 ms ago, cancelled=false, msg=
2016-08-01 15:00:58,982 INFO  [main] client.RpcRetryingCaller: Call exception, tries=27, retries=35, started=390383 ms ago, cancelled=false, msg=
2016-08-01 15:01:19,100 INFO  [main] client.RpcRetryingCaller: Call exception, tries=28, retries=35, started=410501 ms ago, cancelled=false, msg=
2016-08-01 15:01:39,197 INFO  [main] client.RpcRetryingCaller: Call exception, tries=29, retries=35, started=430598 ms ago, cancelled=false, msg=
^C2016-08-01 15:01:54,907 INFO  [Thread-6] client.ConnectionManager$HConnectionImplementation: Closing zookeeper sessionid=0x356353b4f980084
2016-08-01 15:01:54,910 INFO  [Thread-6] zookeeper.ZooKeeper: Session: 0x356353b4f980084 closed
2016-08-01 15:01:54,910 INFO  [main-EventThread] zookeeper.ClientCnxn: EventThread shut down
2016-08-01 15:01:55,221 INFO  [Thread-6] util.HBaseFsck: Finishing hbck



------------------bin/start-hbase.sh실행시 오류내용------------

2016-08-01 14:47:24,312 INFO  [main] master.HMaster: Adding backup master ZNode /hbase/backup-masters/sda1,16000,1470030443501
2016-08-01 14:47:24,354 INFO  [sda1:16000.activeMasterManager] master.ActiveMasterManager: Deleting ZNode for /hbase/backup-masters/sda1,16000,1470030443501 from backup master directory
2016-08-01 14:47:24,356 INFO  [sda1:16000.activeMasterManager] master.ActiveMasterManager: Registered Active Master=sda1,16000,1470030443501
2016-08-01 14:47:24,378 INFO  [master/sda1/XXX.XXX.XXX.43:16000] zookeeper.RecoverableZooKeeper: Process identifier=hconnection-0xea78210 connecting to ZooKeeper ensemble=sda1:2181,sda2:2181,sda3:2181
2016-08-01 14:47:24,379 INFO  [master/sda1/XXX.XXX.XXX.43:16000] zookeeper.ZooKeeper: Initiating client connection, connectString=sda1:2181,sda2:2181,sda3:2181 sessionTimeout=90000 watcher=hconnection-0xea
782100x0, quorum=sda1:2181,sda2:2181,sda3:2181, baseZNode=/hbase
2016-08-01 14:47:24,379 INFO  [master/sda1/XXX.XXX.XXX.43:16000-SendThread(sda3:2181)] zookeeper.ClientCnxn: Opening socket connection to server sda3/XXX.XXX.XXX.31:2181. Will not attempt to authenticate u
sing SASL (unknown error)
2016-08-01 14:47:24,379 INFO  [master/sda1/XXX.XXX.XXX.43:16000-SendThread(sda3:2181)] zookeeper.ClientCnxn: Socket connection established to sda3/XXX.XXX.XXX.31:2181, initiating session
2016-08-01 14:47:24,382 INFO  [master/sda1/XXX.XXX.XXX.43:16000-SendThread(sda3:2181)] zookeeper.ClientCnxn: Session establishment complete on server sda3/XXX.XXX.XXX.31:2181, sessionid = 0x356353b4f98007f
, negotiated timeout = 40000
2016-08-01 14:47:24,395 INFO  [master/sda1/XXX.XXX.XXX.43:16000] regionserver.HRegionServer: ClusterId : 2be4df46-db8b-4fc7-a529-12e571444d54
2016-08-01 14:47:24,436 FATAL [sda1:16000.activeMasterManager] master.HMaster: Failed to become active master
org.apache.hadoop.hbase.util.FileSystemVersionException: HBase file layout needs to be upgraded. You have version null and I want version 8. Consult http://hbase.apache.org/book.html for further informatio
n about upgrading HBase. Is your hbase.rootdir valid? If so, you may need to run 'hbase hbck -fixVersionFile'.
        at org.apache.hadoop.hbase.util.FSUtils.checkVersion(FSUtils.java:677)
        at org.apache.hadoop.hbase.master.MasterFileSystem.checkRootDir(MasterFileSystem.java:455)
        at org.apache.hadoop.hbase.master.MasterFileSystem.createInitialFileSystemLayout(MasterFileSystem.java:146)
        at org.apache.hadoop.hbase.master.MasterFileSystem.<init>(MasterFileSystem.java:126)
        at org.apache.hadoop.hbase.master.HMaster.finishActiveMasterInitialization(HMaster.java:650)
        at org.apache.hadoop.hbase.master.HMaster.access$500(HMaster.java:183)
        at org.apache.hadoop.hbase.master.HMaster$1.run(HMaster.java:1652)
        at java.lang.Thread.run(Thread.java:745)
2016-08-01 14:47:24,438 FATAL [sda1:16000.activeMasterManager] master.HMaster: Unhandled exception. Starting shutdown.
org.apache.hadoop.hbase.util.FileSystemVersionException: HBase file layout needs to be upgraded. You have version null and I want version 8. Consult http://hbase.apache.org/book.html for further informatio
n about upgrading HBase. Is your hbase.rootdir valid? If so, you may need to run 'hbase hbck -fixVersionFile'.
        at org.apache.hadoop.hbase.util.FSUtils.checkVersion(FSUtils.java:677)
        at org.apache.hadoop.hbase.master.MasterFileSystem.checkRootDir(MasterFileSystem.java:455)
        at org.apache.hadoop.hbase.master.MasterFileSystem.createInitialFileSystemLayout(MasterFileSystem.java:146)
        at org.apache.hadoop.hbase.master.MasterFileSystem.<init>(MasterFileSystem.java:126)
        at org.apache.hadoop.hbase.master.HMaster.finishActiveMasterInitialization(HMaster.java:650)
        at org.apache.hadoop.hbase.master.HMaster.access$500(HMaster.java:183)
        at org.apache.hadoop.hbase.master.HMaster$1.run(HMaster.java:1652)
        at java.lang.Thread.run(Thread.java:745)
2016-08-01 14:47:24,438 INFO  [sda1:16000.activeMasterManager] regionserver.HRegionServer: STOPPED: Unhandled exception. Starting shutdown.
2016-08-01 14:47:24,438 INFO  [master/sda1/XXX.XXX.XXX.43:16000] regionserver.HRegionServer: Stopping infoServer
2016-08-01 14:47:24,439 INFO  [master/sda1/XXX.XXX.XXX.43:16000] mortbay.log: Stopped SelectChannelConnector@0.0.0.0:16010
2016-08-01 14:47:24,540 INFO  [master/sda1/XXX.XXX.XXX.43:16000] regionserver.HRegionServer: stopping server sda1,16000,1470030443501
2016-08-01 14:47:24,540 INFO  [master/sda1/XXX.XXX.XXX.43:16000] client.ConnectionManager$HConnectionImplementation: Closing zookeeper sessionid=0x356353b4f98007f
2016-08-01 14:47:24,542 INFO  [master/sda1/XXX.XXX.XXX.43:16000] zookeeper.ZooKeeper: Session: 0x356353b4f98007f closed
2016-08-01 14:47:24,542 INFO  [master/sda1/XXX.XXX.XXX.43:16000-EventThread] zookeeper.ClientCnxn: EventThread shut down
2016-08-01 14:47:24,542 INFO  [master/sda1/XXX.XXX.XXX.43:16000] regionserver.HRegionServer: stopping server sda1,16000,1470030443501; all regions closed.
2016-08-01 14:47:24,542 INFO  [master/sda1/XXX.XXX.XXX.43:16000] hbase.ChoreService: Chore service for: sda1,16000,1470030443501 had [] on shutdown
2016-08-01 14:47:24,545 INFO  [master/sda1/XXX.XXX.XXX.43:16000] ipc.RpcServer: Stopping server on 16000
2016-08-01 14:47:24,545 INFO  [RpcServer.listener,port=16000] ipc.RpcServer: RpcServer.listener,port=16000: stopping
2016-08-01 14:47:24,545 INFO  [RpcServer.responder] ipc.RpcServer: RpcServer.responder: stopped
2016-08-01 14:47:24,545 INFO  [RpcServer.responder] ipc.RpcServer: RpcServer.responder: stopping

번호 제목 날짜 조회 수
741 [Ranger]RangerAdminRESTClient Error gertting pplicies; Received NULL response!!, secureMode=true, user=rangerkms/node01.gooper.com@ GOOPER.COM (auth:KERBEROS), serviceName=cm_kms 2023.06.27 73
740 [vue storefrontui]외부 API통합하기 참고 문서 2022.02.09 80
739 [Encryption Zone]Encryption Zone에 생성된 table을 select할때 HDFS /tmp/zone1에 대한 권한이 없는 경우 2023.06.29 83
738 ./gradlew :composeDown 및 ./gradlew :composeUp 를 성공했을때의 메세지 2023.02.20 84
737 [EncryptionZone]User:testuser not allowed to do "DECRYPT_EEK" on 'testkey' 2023.06.29 89
736 [vi] test.nq파일에서 특정문자열(예, <>)을 찾아서 포함되는 라인을 삭제한 동일한 이름의 파일을 만드는 방법 2017.01.25 98
735 [Impala] alter table구문수행시 "WARNINGS: Impala does not have READ_WRITE access to path 'hdfs://nameservice1/DATA/Temp/DB/source/table01_ccd'" 발생시 조치 2024.04.26 98
734 CM의 Impala->Query tab에서 FINISHED query가 보이지 않는 현상 2021.08.31 99
733 restaurant-controller,에서 등록 예시 2022.04.30 99
732 주문히스토리 조회 2022.04.30 99
731 [Hue metadata]Oracle에 있는 Hue 메타정보 테이블을 이용하여 coordinator와 workflow관계 목록을 추출하는 방법 2023.08.22 99
730 [Cloudera Agent] Metadata-Plugin throttling_logger INFO (713 skipped) Unable to send data to nav server. Will try again. 2022.05.16 103
729 oozie의 sqoop action수행시 ooize:launcher의 applicationId를 이용하여 oozie:action의 applicationId및 관련 로그를 찾는 방법 2023.07.26 104
728 [CDP7.1.6,HDFS]HDFS파일을 삭제하고 Trash비움이 완료된후에도 HDFS 공간을 차지하고 있는 경우 확인/조치 방법 2023.07.17 107
727 [CDP7.1.7, Replication]Encryption Zone내 HDFS파일을 비Encryption Zone으로 HDFS Replication시 User hdfs가 아닌 hadoop으로 수행하는 방법 2024.01.15 110
726 주문 생성 데이터 예시 2022.04.30 112
725 호출 url현황 2023.02.21 112
724 [CDP7.1.7, Hive Replication]Hive Replication진행중 "The following columns have types incompatible with the existing columns in their respective positions " 오류 2023.12.27 116
723 eclipse 3.1 단축키 정리파일 2017.01.02 118
722 [CDP7.1.7]Oozie job에서 ERROR: Kudu error(s) reported, first error: Timed out: Failed to write batch of 774 ops to tablet 8003f9a064bf4be5890a178439b2ba91가 발생하면서 쿼리가 실패하는 경우 2024.01.05 118
위로