No data nodes are started

2019-01-16 10:22发布


I am trying to setup Hadoop version in a pseudo distributed configuration using the following guide:

After running the script I run "jps".

I get this output:

4825 NameNode
5391 TaskTracker
5242 JobTracker
5477 Jps
5140 SecondaryNameNode

When I try to add information to the hdfs using:

bin/hadoop fs -put conf input

I got an error:

hadoop@m1a2:~/software/hadoop$ bin/hadoop fs -put conf input
12/04/10 18:15:31 WARN hdfs.DFSClient: DataStreamer Exception: org.apache.hadoop.ipc.RemoteException: File /user/hadoop/input/core-site.xml could only be replicated to 0 nodes, instead of 1
        at org.apache.hadoop.hdfs.server.namenode.FSNamesystem.getAdditionalBlock(
        at org.apache.hadoop.hdfs.server.namenode.NameNode.addBlock(
        at sun.reflect.GeneratedMethodAccessor6.invoke(Unknown Source)
        at sun.reflect.DelegatingMethodAccessorImpl.invoke(
        at java.lang.reflect.Method.invoke(
        at org.apache.hadoop.ipc.RPC$
        at org.apache.hadoop.ipc.Server$Handler$
        at org.apache.hadoop.ipc.Server$Handler$
        at Method)
        at org.apache.hadoop.ipc.Server$

        at org.apache.hadoop.ipc.RPC$Invoker.invoke(
        at $Proxy1.addBlock(Unknown Source)
        at sun.reflect.NativeMethodAccessorImpl.invoke0(Native Method)
        at sun.reflect.NativeMethodAccessorImpl.invoke(
        at sun.reflect.DelegatingMethodAccessorImpl.invoke(
        at java.lang.reflect.Method.invoke(
        at $Proxy1.addBlock(Unknown Source)
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream.locateFollowingBlock(
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream.nextBlockOutputStream(
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream.access$2000(
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream$

12/04/10 18:15:31 WARN hdfs.DFSClient: Error Recovery for block null bad datanode[0] nodes == null
12/04/10 18:15:31 WARN hdfs.DFSClient: Could not get block locations. Source file "/user/hadoop/input/core-site.xml" - Aborting...
put: File /user/hadoop/input/core-site.xml could only be replicated to 0 nodes, instead of 1
12/04/10 18:15:31 ERROR hdfs.DFSClient: Exception closing file /user/hadoop/input/core-site.xml : org.apache.hadoop.ipc.RemoteException: File /user/hadoop/input/core-site.xml could only be replicated to 0 nodes, instead of 1
        at org.apache.hadoop.hdfs.server.namenode.FSNamesystem.getAdditionalBlock(
        at org.apache.hadoop.hdfs.server.namenode.NameNode.addBlock(
        at sun.reflect.GeneratedMethodAccessor6.invoke(Unknown Source)
        at sun.reflect.DelegatingMethodAccessorImpl.invoke(
        at java.lang.reflect.Method.invoke(
        at org.apache.hadoop.ipc.RPC$
        at org.apache.hadoop.ipc.Server$Handler$
        at org.apache.hadoop.ipc.Server$Handler$
        at Method)
        at org.apache.hadoop.ipc.Server$

org.apache.hadoop.ipc.RemoteException: File /user/hadoop/input/core-site.xml could only be replicated to 0 nodes, instead of 1
        at org.apache.hadoop.hdfs.server.namenode.FSNamesystem.getAdditionalBlock(
        at org.apache.hadoop.hdfs.server.namenode.NameNode.addBlock(
        at sun.reflect.GeneratedMethodAccessor6.invoke(Unknown Source)
        at sun.reflect.DelegatingMethodAccessorImpl.invoke(
        at java.lang.reflect.Method.invoke(
        at org.apache.hadoop.ipc.RPC$
        at org.apache.hadoop.ipc.Server$Handler$
        at org.apache.hadoop.ipc.Server$Handler$
        at Method)
        at org.apache.hadoop.ipc.Server$

        at org.apache.hadoop.ipc.RPC$Invoker.invoke(
        at $Proxy1.addBlock(Unknown Source)
        at sun.reflect.NativeMethodAccessorImpl.invoke0(Native Method)
        at sun.reflect.NativeMethodAccessorImpl.invoke(
        at sun.reflect.DelegatingMethodAccessorImpl.invoke(
        at java.lang.reflect.Method.invoke(
        at $Proxy1.addBlock(Unknown Source)
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream.locateFollowingBlock(
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream.nextBlockOutputStream(
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream.access$2000(
        at org.apache.hadoop.hdfs.DFSClient$DFSOutputStream$

I am not totally sure but I believe that this may have to do with the fact that the datanode is not running.

Does anybody know what I have done wrong, or how to fix this problem?

EDIT: This is the datanode.log file:

2012-04-11 12:27:28,977 INFO org.apache.hadoop.hdfs.server.datanode.DataNode: STARTUP_MSG:
STARTUP_MSG: Starting DataNode
STARTUP_MSG:   host = m1a2/
STARTUP_MSG:   args = []
STARTUP_MSG:   version =
STARTUP_MSG:   build = -r 1099333; compiled by 'oom' on Wed May  4 07:57:50 PDT 2011
2012-04-11 12:27:29,166 INFO org.apache.hadoop.metrics2.impl.MetricsConfig: loaded properties from
2012-04-11 12:27:29,181 INFO org.apache.hadoop.metrics2.impl.MetricsSourceAdapter: MBean for source MetricsSystem,sub=Stats registered.
2012-04-11 12:27:29,183 INFO org.apache.hadoop.metrics2.impl.MetricsSystemImpl: Scheduled snapshot period at 10 second(s).
2012-04-11 12:27:29,183 INFO org.apache.hadoop.metrics2.impl.MetricsSystemImpl: DataNode metrics system started
2012-04-11 12:27:29,342 INFO org.apache.hadoop.metrics2.impl.MetricsSourceAdapter: MBean for source ugi registered.
2012-04-11 12:27:29,347 WARN org.apache.hadoop.metrics2.impl.MetricsSystemImpl: Source name ugi already exists!
2012-04-11 12:27:29,615 ERROR org.apache.hadoop.hdfs.server.datanode.DataNode: Incompatible namespaceIDs in /tmp/hadoop-hadoop/dfs/data: namenode namespaceID = 301052954; datanode namespaceID = 229562149
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.doTransition(
        at org.apache.hadoop.hdfs.server.datanode.DataStorage.recoverTransitionRead(
        at org.apache.hadoop.hdfs.server.datanode.DataNode.startDataNode(
        at org.apache.hadoop.hdfs.server.datanode.DataNode.<init>(
        at org.apache.hadoop.hdfs.server.datanode.DataNode.makeInstance(
        at org.apache.hadoop.hdfs.server.datanode.DataNode.instantiateDataNode(
        at org.apache.hadoop.hdfs.server.datanode.DataNode.createDataNode(
        at org.apache.hadoop.hdfs.server.datanode.DataNode.secureMain(
        at org.apache.hadoop.hdfs.server.datanode.DataNode.main(

2012-04-11 12:27:29,617 INFO org.apache.hadoop.hdfs.server.datanode.DataNode: SHUTDOWN_MSG:
SHUTDOWN_MSG: Shutting down DataNode at m1a2/


That error you are getting in the DN log is described here:

From that page:

At the moment, there seem to be two workarounds as described below.

Workaround 1: Start from scratch

I can testify that the following steps solve this error, but the side effects won’t make you happy (me neither). The crude workaround I have found is to:

  1. Stop the cluster
  2. Delete the data directory on the problematic DataNode: the directory is specified by in conf/hdfs-site.xml; if you followed this tutorial, the relevant directory is /app/hadoop/tmp/dfs/data
  3. Reformat the NameNode (NOTE: all HDFS data is lost during this process!)
  4. Restart the cluster

When deleting all the HDFS data and starting from scratch does not sound like a good idea (it might be ok during the initial setup/testing), you might give the second approach a try.

Workaround 2: Updating namespaceID of problematic DataNodes

Big thanks to Jared Stehler for the following suggestion. I have not tested it myself yet, but feel free to try it out and send me your feedback. This workaround is “minimally invasive” as you only have to edit one file on the problematic DataNodes:

  1. Stop the DataNode
  2. Edit the value of namespaceID in /current/VERSION to match the value of the current NameNode
  3. Restart the DataNode

If you followed the instructions in my tutorials, the full path of the relevant files are:

NameNode: /app/hadoop/tmp/dfs/name/current/VERSION

DataNode: /app/hadoop/tmp/dfs/data/current/VERSION

(background: is by default set to

${hadoop.tmp.dir}/dfs/data, and we set hadoop.tmp.dir

in this tutorial to /app/hadoop/tmp).

If you wonder how the contents of VERSION look like, here’s one of mine:

# contents of /current/VERSION







Okay, I post this once more:

In case someone needs this, for newer version of Hadoop (basically I am running 2.4.0)

  • In this case stop the cluster sbin/

  • Then go to /etc/hadoop for config files.

In the file: hdfs-site.xml Look out for directory paths corresponding to

  • Delete both the directories recursively (rm -r).

  • Now format the namenode via bin/hadoop namenode -format

  • And finally sbin/

Hope this helps.


I had the same issue on pseudo node using hadoop1.1.2 So I ran bin/ to stop the cluster then saw the configuration of my hadoop tmp directory in hdfs-site.xml


So I went into /root/data/hdfstmp and deleted all the files using command (you may loose ur data)

rm -rf *

and then format namenode again

bin/hadoop namenode -format

and then start the cluster using


Main reason is bin/hadoop namenode -format didn't remove the old data. So we have to delete it manually.


Do following steps:

1. bin/
2. remove dfs/ and mapred/ folder of hadoop.tmp.dir in core-site.xml
3. bin/hadoop namenode -format
4. bin/
5. jps


Try formatting your datanode and restart it.


I have been using CDH4 as my version of hadoop and have been having trouble configuring it. Even after trying to reformat my namenode I was still receiving the error.

My VERSION file was located in


You can find the location of the HDFS cache directory by looking for the hadoop.tmp.dir property:

more /etc/hadoop/conf/hdfs-site.xml 

I found that by doing

cd /var/lib/hadoop-hdfs/cache/
rm -rf *

and then reformatting the namenode I was finally able to fix the issue. Thanks to the first reply for helping me to figure out what folder I needed to bomb.


I tried with the approach 2 as suggested by Jared Stehler in the Chris Shain answer and i can confirm that after doing these changes , i was able to resolve the above mentioned problem.

I used the same version number for both the name and data VERSION file. Mean to say copied the version number from file VERSION inside (/app/hadoop/tmp/dfs/name/current) to VERSION inside (/app/hadoop/tmp/dfs/data/current) and it worked like charm

Cheers !


I encountered this issue when using an unmodified cloudera quickstart vm 4.4.0-1

For reference, the cloudera manager said my datanode was in good health, even though the error message in the DataStreamer stacktrace said no datanodes were running.

credit goes to workaround #2 from but i'll detail my specific experience using the cloudera quickstart vm.

Specifically, I did:
in this order, stop the services hue1, hive1, mapreduce1, hdfs1 via the cloudera manager http://localhost.localdomain:7180/cmf/services/status

found my VERSION files via:
sudo find / -name VERSION

i got:


i checked the contents of those files, but they all had a matching namespaceID except one file was just totally missing it. so i added an entry to it.

then i restarted the services in reverse order via the cloudera manager. now i can -put stuff onto hdfs.


In my case, I wrongly set one destination for and The correct format is




I have the same problem with datanode missing and i follow this step that worked for me

1.find the folder that datanode located in. cd hadoop/hadoopdata/hdfs 2.look in the folder and you will see what file you have in hdfs ls 3.delete the datanode folder because it is old version of datanode rm -rf/datanode/* 4. you will get the new version after run the previous command 5. start new datanode start datanode 6. refresh the web services. You will see the lost node appears my terminal

标签: hadoop hdfs