How to run the example 'Run mapreduce on a Hadoop Cluster'?

Question

Jingyu Ru 2015 年 7 月 22 日

0
リンク

この質問への直接リンク

https://jp.mathworks.com/matlabcentral/answers/230805-how-to-run-the-example-run-mapreduce-on-a-hadoop-cluster

コメント済み: lov kumar 2019 年 6 月 4 日

採用された回答: Esther

MATLAB Online で開く

Hadoop version 1.2.1 Matlab version 2015a

Linux ubuntu 14.

I install Hadoop for a pseudo-distributed configuration (all of the Hadoop daemons run on a single host).

It is success to run the example 'wordcount'in Hadoop.

And it is success to read the data from the HDFS through the Matlab.

But when I try to run the example in Matlab 'Run mapreduce on a Hadoop Cluster',I failed.

It shows that the Map 0% and Reduce 0%.

There is my Matlab codes

setenv('HADOOP_HOME','/home/rjy/soft/hadoop-1.2.1');

cluster = parallel.cluster.Hadoop;

cluster.HadoopProperties('mapred.job.tracker') = 'localhost:50031';

cluster.HadoopProperties('fs.default.name') = 'hdfs://localhost:8020';

% ds = datastore('hdfs:/user/root/rrr')

outputFolder = '/home/rjy/logs/hadooplog';

mr = mapreducer(cluster);

ds = datastore('airlinesmall.csv','TreatAsMissing','NA',...

'SelectedVariableNames','ArrDelay','ReadSize',1000);

preview(ds)

meanDelay = mapreduce(ds,@meanArrivalDelayMapper,@meanArrivalDelayReducer,mr,...

'OutputFolder',outputFolder)

Error log is shown as follows,

Error using mapreduce (line 100)

The HADOOP job failed to complete.

Error in run_mapreduce_on_a_hadoop (line 12)

meanDelay = mapreduce(ds,@meanArrivalDelayMapper,@meanArrivalDelayReducer,mr,...

Caused by:

    The HADOOP job was not able to start MATLAB for attempt 1 of 'MAP' task 0. The user
    home directory '/homes/' for the cluster either did not exist or was not writable by
    the HADOOP user. Check the documentation on how to set the user home directory for the
    cluster.

What's the '/homes/' means? I have never seen it before. My hadoop directory is './home/rjy/soft/hadoop-1.2.1'

I wish to know how to solve this problem. I have tried many methods to solve it.

Please give me some suggestions. Thanks.

0 件のコメント
-2 件の古いコメントを表示-2 件の古いコメントを非表示

サインインしてコメントする。

サインインしてこの質問に回答する。

Answer 1

Esther 2015 年 7 月 26 日

2
リンク

この回答への直接リンク

https://jp.mathworks.com/matlabcentral/answers/230805-how-to-run-the-example-run-mapreduce-on-a-hadoop-cluster#answer_187215

Hi Jingyu,

Try editing the mapred-site.xml to set the value of this property: mapreduce.admin.user.home.dir

which will be set to a value of that of your hadoop directory: './home/rjy/soft/hadoop-1.2.1'

Maybe this will force it to look to your hadoop installation instead of /homes.

Referencing: http://www.mathworks.com/matlabcentral/answers/82060-mcr-problems-with-linux-hadoop

Hope this can solve the problem.

Cheers, Esther

2 件のコメント
なしを表示なしを非表示

Jingyu Ru 2015 年 7 月 28 日

Thanks a lot, I have solved the problem by set the mapred-site.xml,

property name>mapreduce.admin.user.home.dir</name value/home/rjy/soft/hadoop-1.2.1</value> /property /configuration

Thank you very much!!! I'm so exciting!!!

Esther 2015 年 7 月 31 日

You're welcome!

サインインしてコメントする。

Answer 2

lov kumar 2019 年 5 月 28 日

0
リンク

この回答への直接リンク

https://jp.mathworks.com/matlabcentral/answers/230805-how-to-run-the-example-run-mapreduce-on-a-hadoop-cluster#answer_376864

How to fix this error

Error using mapreduce (line 124)

The HADOOP job failed to submit. It is possible that there is some issue with the HADOOP configuration.

Error in bg (line 14)

meanDelay = mapreduce(ds,@meanArrivalDelayMapper,@meanArrivalDelayReducer,mr,...

4 件のコメント
2 件の古いコメントを表示2 件の古いコメントを非表示

lov kumar 2019 年 6 月 4 日

I am using this code:

setenv('HADOOP_PREFIX','C:/hadoop-2.8.0');

getenv('HADOOP_PREFIX');

ds = datastore('hdfs://localhost:9000/lov/airlinesmall.csv',...

'TreatAsMissing','NA',...

'SelectedVariableNames',{'UniqueCarrier','ArrDelay'});

cluster = parallel.cluster.Hadoop;

cluster.HadoopProperties('mapred.job.tracker') = 'localhost:9000';

cluster.HadoopProperties('fs.default.name') = 'hdfs://localhost:8088';

disp(cluster);

mapreducer(cluster);

result = mapreduce(ds,@maxArrivalDelayMapper,@maxArrivalDelayReducer,'OutputFolder','hdfs://localhost:9000/lov');

lov kumar 2019 年 6 月 4 日

I am using single machine.

サインインしてコメントする。

How to run the example 'Run mapreduce on a Hadoop Cluster'?

0 件のコメント
-2 件の古いコメントを表示-2 件の古いコメントを非表示

採用された回答

2 件のコメント
なしを表示なしを非表示

その他の回答 (1 件)

4 件のコメント
2 件の古いコメントを表示2 件の古いコメントを非表示

参考

カテゴリ

タグ

製品

Community Treasure Hunt

How to run the example 'Run mapreduce on a Hadoop Cluster'?

0 件のコメント -2 件の古いコメントを表示-2 件の古いコメントを非表示

採用された回答

2 件のコメント なしを表示なしを非表示

その他の回答 (1 件)

4 件のコメント 2 件の古いコメントを表示2 件の古いコメントを非表示

参考

カテゴリ

タグ

製品

Community Treasure Hunt

0 件のコメント
-2 件の古いコメントを表示-2 件の古いコメントを非表示

2 件のコメント
なしを表示なしを非表示

4 件のコメント
2 件の古いコメントを表示2 件の古いコメントを非表示