Java hdfs namenode和datanode未基于dfs.replication值启动

Java hdfs namenode和datanode未基于dfs.replication值启动,java,hadoop,hdfs,yarn,Java,Hadoop,Hdfs,Yarn,由于我正处于Hadoop的学习阶段,我在Hadoop单集群设置方面遇到了问题。我正在使用Hadoop2.9.0和Java8。我已经完成了设置,如下所示 核心站点.xml <configuration> <property> <name>hadoop.tmp.dir</name> <value>/home/ubuntu/hadooptmp/hadoop-${user.name}</value> <description&

由于我正处于Hadoop的学习阶段,我在Hadoop单集群设置方面遇到了问题。我正在使用Hadoop2.9.0和Java8。我已经完成了设置,如下所示

核心站点.xml

<configuration>
<property>
<name>hadoop.tmp.dir</name>
<value>/home/ubuntu/hadooptmp/hadoop-${user.name}</value>
<description>A base for other temporary directories.</description>
</property>
<property>
<name>fs.default.name</name>
<value>hdfs://localhost:9000</value>
</property>
</configuration>
<configuration>
   <property>
      <name>mapreduce.framework.name</name>
      <value>localhost:9001</value>
   </property>
</configuration>
<configuration>
<!-- Site specific YARN configuration properties -->
<property>
<name>yarn.nodemanager.aux-services</name>
<value>mapreduce_shuffle</value>
</property>
</configuration>
<configuration>
<property>
<name>dfs.replication</name>
<value>1</value>
</property>
<property><name>dfs.name.dir</name>
<value>file:///home/ubuntu/hadoop/hdfs/namenode</value>
<final>true</final>
</property>
<property>
<name>dfs.data.dir</name>
<value>file:///home/ubuntu/hadoop/hdfs/datanode</value>
<final>true</final>
</property>
</configuration>
现在我有stop-all.sh,如果我将hdfs-site.xml中的dfs.replication的值更改为0(有些人将此称为解决方案),然后再次启动start-all.sh。身份是-

9024 DataNode
9362 ResourceManager
9493 NodeManager
9815 Jps
正如我们在本例中看到的,Name节点已停止运行。我尝试过格式化这两种格式,但都没有成功。我不明白我在安装中做错了什么,无法同时运行Namenode和datanode

我已经在aws ubuntu服务器上完成了所有的单集群设置。我曾尝试将mapreduce.framework.name在mapred-site.xml中的值命名为“纱线”,但没有帮助。 请让我知道是否有解决方案

下面添加了Datanode日志

2018-08-23 15:08:14,494 INFO org.apache.hadoop.hdfs.server.common.Storage: Using 1 threads to upgrade data directories (dfs.datanode.parallel.volumes.load.threads.num=1, dataDirs=1)
2018-08-23 15:08:14,506 INFO org.apache.hadoop.hdfs.server.common.Storage: Lock on /home/ubuntu/hadoop/hdfs/datanode/in_use.lock acquired by nodename 1827@ip-172-31-29-76.us-east-2.compute.internal
2018-08-23 15:08:14,509 WARN org.apache.hadoop.hdfs.server.common.Storage: Failed to add storage directory [DISK]file:/home/ubuntu/hadoop/hdfs/datanode/
java.io.IOException: Incompatible clusterIDs in /home/ubuntu/hadoop/hdfs/datanode: namenode clusterID = CID-e5d9317d-7c4d-40eb-8fa2-030c5a8cfad6; datanode clusterID = CID-f416763a-ed6f-41d7-8b9d-12c298c5d779
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.doTransition(DataStorage.java:760)
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.loadStorageDirectory(DataStorage.java:293)
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.loadDataStorage(DataStorage.java:409)
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.addStorageLocations(DataStorage.java:388)
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.recoverTransitionRead(DataStorage.java:556)
        at org.apache.hadoop.hdfs.server.datanode.DataNode.initStorage(DataNode.java:1649)
        at org.apache.hadoop.hdfs.server.datanode.DataNode.initBlockPool(DataNode.java:1610)
        at org.apache.hadoop.hdfs.server.datanode.BPOfferService.verifyAndSetNamespaceInfo(BPOfferService.java:374)
        at org.apache.hadoop.hdfs.server.datanode.BPServiceActor.connectToNNAndHandshake(BPServiceActor.java:280)
        at org.apache.hadoop.hdfs.server.datanode.BPServiceActor.run(BPServiceActor.java:816)
        at java.lang.Thread.run(Thread.java:748)
2018-08-23 15:08:14,516 ERROR org.apache.hadoop.hdfs.server.datanode.DataNode: Initialization failed for Block pool <registering> (Datanode Uuid 85d94537-7ccc-4d7d-aa2c-48c65a453399) service to localhost/127.0.0.1:9000. Exiting.
java.io.IOException: All specified directories have failed to load.
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.recoverTransitionRead(DataStorage.java:557)
        at org.apache.hadoop.hdfs.server.datanode.DataNode.initStorage(DataNode.java:1649)
        at org.apache.hadoop.hdfs.server.datanode.DataNode.initBlockPool(DataNode.java:1610)
        at org.apache.hadoop.hdfs.server.datanode.BPOfferService.verifyAndSetNamespaceInfo(BPOfferService.java:374)
        at org.apache.hadoop.hdfs.server.datanode.BPServiceActor.connectToNNAndHandshake(BPServiceActor.java:280)
        at org.apache.hadoop.hdfs.server.datanode.BPServiceActor.run(BPServiceActor.java:816)
        at java.lang.Thread.run(Thread.java:748)
2018-08-23 15:08:14,516 WARN org.apache.hadoop.hdfs.server.datanode.DataNode: Ending block pool service for: Block pool <registering> (Datanode Uuid 85d94537-7ccc-4d7d-aa2c-48c65a453399) service to localhost/127.0.0.1:9000
2018-08-23 15:08:14,518 INFO org.apache.hadoop.hdfs.server.datanode.DataNode: Removed Block pool <registering> (Datanode Uuid 85d94537-7ccc-4d7d-aa2c-48c65a453399)
2018-08-23 15:08:16,520 WARN org.apache.hadoop.hdfs.server.datanode.DataNode: Exiting Datanode
2018-08-23 15:08:16,531 INFO org.apache.hadoop.hdfs.server.datanode.DataNode: SHUTDOWN_MSG:
/************************************************************
SHUTDOWN_MSG: Shutting down DataNode at ip-172-31-29-76.us-east-2.compute.internal/172.31.29.76
2018-08-23 15:08:14494 INFO org.apache.hadoop.hdfs.server.common.Storage:使用1个线程升级数据目录(dfs.datanode.parallel.volumes.load.threads.num=1,dataDirs=1)
2018-08-23 15:08:14506 INFO org.apache.hadoop.hdfs.server.common.Storage:Lock on/home/ubuntu/hadoop/hdfs/datanode/in_use.Lock被nodename收购1827@ip-172-31-29-76.us-east-2.compute.internal
2018-08-23 15:08:14509警告org.apache.hadoop.hdfs.server.common.Storage:未能添加存储目录[磁盘]文件:/home/ubuntu/hadoop/hdfs/datanode/
java.io.IOException:home/ubuntu/hadoop/hdfs/datanode中不兼容的clusterID:namenode clusterID=CID-e5d9317d-7c4d-40eb-8fa2-030c5a8cfad6;数据节点群集ID=CID-f416763a-ed6f-41d7-8b9d-12c298c5d779
位于org.apache.hadoop.hdfs.server.datanode.DataStorage.doTransition(DataStorage.java:760)
位于org.apache.hadoop.hdfs.server.datanode.DataStorage.loadStorageDirectory(DataStorage.java:293)
位于org.apache.hadoop.hdfs.server.datanode.DataStorage.loadDataStorage(DataStorage.java:409)
位于org.apache.hadoop.hdfs.server.datanode.DataStorage.addStorageLocations(DataStorage.java:388)
位于org.apache.hadoop.hdfs.server.datanode.DataStorage.recoverTransitionRead(DataStorage.java:556)
位于org.apache.hadoop.hdfs.server.datanode.datanode.initStorage(datanode.java:1649)
位于org.apache.hadoop.hdfs.server.datanode.datanode.initBlockPool(datanode.java:1610)
位于org.apache.hadoop.hdfs.server.datanode.BPOfferService.verifyAndSetNamespaceInfo(BPOfferService.java:374)
位于org.apache.hadoop.hdfs.server.datanode.BPServiceActor.connecttonandhandshake(BPServiceActor.java:280)
位于org.apache.hadoop.hdfs.server.datanode.BPServiceActor.run(BPServiceActor.java:816)
运行(Thread.java:748)
2018-08-23 15:08:14516错误org.apache.hadoop.hdfs.server.datanode.datanode:localhost/127.0.0.1:9000的块池(datanode Uuid 85d94537-7ccc-4d7d-aa2c-48c65a453399)服务初始化失败。退出。
java.io.IOException:无法加载所有指定的目录。
位于org.apache.hadoop.hdfs.server.datanode.DataStorage.recoverTransitionRead(DataStorage.java:557)
位于org.apache.hadoop.hdfs.server.datanode.datanode.initStorage(datanode.java:1649)
位于org.apache.hadoop.hdfs.server.datanode.datanode.initBlockPool(datanode.java:1610)
位于org.apache.hadoop.hdfs.server.datanode.BPOfferService.verifyAndSetNamespaceInfo(BPOfferService.java:374)
位于org.apache.hadoop.hdfs.server.datanode.BPServiceActor.connecttonandhandshake(BPServiceActor.java:280)
位于org.apache.hadoop.hdfs.server.datanode.BPServiceActor.run(BPServiceActor.java:816)
运行(Thread.java:748)
2018-08-23 15:08:14516警告org.apache.hadoop.hdfs.server.datanode.datanode:终止块池服务,用于:本地主机的块池(数据节点Uuid 85d94537-7ccc-4d7d-aa2c-48c65a453399)服务/127.0.0.1:9000
2018-08-23 15:08:14518 INFO org.apache.hadoop.hdfs.server.datanode.datanode:删除的块池(datanode Uuid 85d94537-7ccc-4d7d-aa2c-48c65a453399)
2018-08-23 15:08:16520警告org.apache.hadoop.hdfs.server.datanode.datanode:正在退出datanode
2018-08-23 15:08:16531 INFO org.apache.hadoop.hdfs.server.datanode.datanode:SHUTDOWN\u MSG:
/************************************************************
关机消息:正在关闭ip-172-31-29-76.us-east-2.compute.internal/172.31.29.76处的数据节点

我认为问题在于name节点和datanode的集群ID之间的冲突。我运行了一个命令“sudorm-rf/home/ubuntu/hadoop/hdfs/datanode/*”,修复了Namenode和datanode都在运行的问题

4632 NameNode
4793 DataNode
5145 ResourceManager
5531 Jps
5276 NodeManager
4989 SecondaryNameNode

谢谢你们的帮助。如果我遗漏了什么,请告诉我。

dfs.replication as 0不是解决方案,请将你的dfs.replication as 1更改为start-all.sh,如果datanode没有启动,请共享datanode的日志,这些日志应该会为问题提供一些指针。如果你能显示服务的日志,那就太好了。您正在格式化namenode吗?Hdfs也不应设置为localhost。MapReduce框架也需要改进,它不是一个服务器地址。。。这是由warn-site.xml定义的。请查找问题中添加的Datanode日志。希望这有帮助。
4632 NameNode
4793 DataNode
5145 ResourceManager
5531 Jps
5276 NodeManager
4989 SecondaryNameNode