Hadoop & Spark & Hive & HBase

Hadoop:
http://hadoop.apache.org/docs/r2.6.4/hadoop-project-dist/hadoop-common/SingleCluster.html



bin/hdfs namenode -format
sbin/start-dfs.sh


 http://localhost:50070/
 



bin/hdfs dfs -mkdir /user
bin/hdfs dfs -mkdir /user/<username>


these are for testing:


bin/hdfs dfs -put etc/hadoop input
bin/hadoop jar share/hadoop/mapreduce/hadoop-mapreduce-examples-2.6.4.jar grep input output 'dfs[a-z.]+'
bin/hdfs dfs -cat output/*


testing results:


6       dfs.audit.logger
4       dfs.class
3       dfs.server.namenode.
2       dfs.period
2       dfs.audit.log.maxfilesize
2       dfs.audit.log.maxbackupindex
1       dfsmetrics.log
1       dfsadmin
1       dfs.servers
1       dfs.replication
1       dfs.file







YARN: 
ResourceManager

./sbin/start-yarn.sh


Http://localhost:8088/
 



HistoryServer



./sbin/mr-jobhistory-daemon.sh start historyserver





http://localhost:19888/
 




Spark:




http://spark.apache.org/docs/1.6.2/
  start: 


./sbin/start-master.sh


 http://localhost:8080/


 start worker:




./sbin/start-slaves.sh spark://<your-computer-name>:7077  


You will see:


Alive Workers: 1

 http://localhost:8080/



This is for testing:





./bin/spark-shell --master spark://<your-computer-name>:7077





You will see the scala shell.
use :q to quit.

To see the history:

http://spark.apache.org/docs/latest/monitoring.html

http://blog.chinaunix.net/uid-29454152-id-5641909.html

http://www.cloudera.com/documentation/cdh/5-1-x/CDH5-Installation-Guide/cdh5ig_spark_configure.html

./sbin/start-history-server.sh

http://localhost:18080/

Hive:

https://cwiki.apache.org/confluence/display/Hive/GettingStarted

https://cwiki.apache.org/confluence/display/Hive/Setting+Up+HiveServer2

http://www.360doc.com/content/16/0411/19/2795334_549791350.shtml

Bug:

in mysql 5.7 you should use :

jdbc:mysql://localhost:3306/hivedb?useSSL=false&amp;createDatabaseIfNotExist=true

start hiveserver2:

 nohup hiveserver2 &

http://localhost:10002/

Bug：

User:  is not allowed to impersonate anonymous (state=,code=0)

https://community.hortonworks.com/questions/42483/user-hive-is-not-allowed-to-impersonate-anonymous.html

http://stackoverflow.com/questions/31228420/how-to-run-hive-on-spark-job-from-beeline-or-any-jdbc-client

See more:

https://hadoop.apache.org/docs/current/hadoop-project-dist/hadoop-common/Superusers.html

Hwi 界面Bug：

HWI WAR file not found at

pack the war file yourself, then copy it to the right place, then add needed setting into hive-site.xml

http://blog.csdn.net/gao634209276/article/details/51426371

http://blog.csdn.net/duguduchong/article/details/8852425


Problem: failed to create task or type componentdef
Or:
Could not create task or type of type: componentdef

sudo apt-get install libjasperreports-java

sudo apt-get install ant

_________________________________________________________________________________not finished

自定义配置：

http://blog.csdn.net/reesun/article/details/8556078

数据库连接软件：

默认用户名就是登录账号密码为空

http://blog.sina.com.cn/s/blog_76923bd80102wi3g.html

语法

https://cwiki.apache.org/confluence/display/Hive/LanguageManual+DDL

more info:

http://stackoverflow.com/questions/35476468/what-is-the-difference-between-the-hive-metastore-in-derby-vs-the-one-in-hive-wa

http://www.2cto.com/database/201408/325554.html

HBase

http://hbase.apache.org/book.html#quickstart

./bin/start-hbase.sh

http://localhost:16010/

HBase & Hive

Hive & Shark & SparkSQL

Spark SQL架构如下图所示:

http://blog.csdn.net/wzy0623/article/details/52249187

http://lib.csdn.net/article/spark/33925

phoenix

queryserver.py start

jdbc:phoenix:thin:url=http://localhost:8765;serialization=PROTOBUF

Or:

phoenix-sqlline.py localhost:2181

来自为知笔记(Wiz)

Hadoop & Spark & Hive & HBase的更多相关文章

大数据学习系列之七 ----- Hadoop+Spark+Zookeeper+HBase+Hive集群搭建图文详解
引言在之前的大数据学习系列中,搭建了Hadoop+Spark+HBase+Hive 环境以及一些测试.其实要说的话,我开始学习大数据的时候,搭建的就是集群,并不是单机模式和伪分布式.至于为什么先写单 ...
HADOOP+SPARK+ZOOKEEPER+HBASE+HIVE集群搭建(转)
原文地址:https://www.cnblogs.com/hanzhi/articles/8794984.html 目录引言目录一环境选择 1集群机器安装图 2配置说明 3下载地址二集群的相关 ...
hadoop之hive&hbase互操作
大家都知道,hive的SQL操作非常方便,但是查询过程中需要启动MapReduce,无法做到实时响应. hbase是hadoop家族中的分布式数据库,与传统关系数据库不同,它底层采用列存储格式,扩展性 ...
大数据学习系列之九---- Hive整合Spark和HBase以及相关测试
前言在之前的大数据学习系列之七 ----- Hadoop+Spark+Zookeeper+HBase+Hive集群搭建中介绍了集群的环境搭建,但是在使用hive进行数据查询的时候会非常的慢,因为h ...
Hadoop + Hive + HBase + Kylin伪分布式安装
问题导读 1. Centos7如何安装配置? 2. linux网络配置如何进行? 3. linux环境下java 如何安装? 4. linux环境下SSH免密码登录如何配置? 5. linux环境下H ...
【原创】大叔问题定位分享（16）spark写数据到hive外部表报错ClassCastException: org.apache.hadoop.hive.hbase.HiveHBaseTableOutputFormat cannot be cast to org.apache.hadoop.hive.ql.io.HiveOutputFormat
spark 2.1.1 spark在写数据到hive外部表(底层数据在hbase中)时会报错 Caused by: java.lang.ClassCastException: org.apache.h ...
Docker搭建大数据集群 Hadoop Spark HBase Hive Zookeeper Scala
Docker搭建大数据集群给出一个完全分布式hadoop+spark集群搭建完整文档,从环境准备(包括机器名,ip映射步骤,ssh免密,Java等)开始,包括zookeeper,hadoop,hiv ...
大数据技术生态圈形象比喻（Hadoop、Hive、Spark 关系）
[摘要] 知乎上一篇很不错的科普文章,介绍大数据技术生态圈(Hadoop.Hive.Spark )的关系. 链接地址:https://www.zhihu.com/question/27974418 [ ...
spark读取hbase形成RDD，存入hive或者spark_sql分析
object SaprkReadHbase { var total:Int = 0 def main(args: Array[String]) { val spark = SparkSession . ...

随机推荐

WC2019退役记
sb题不会,暴力写不完,被全场吊着打,AFO
【性能测试】：JVM内存监控策略的方法，以及监控结果说明
JVM内存监控主要在稳定性压测期间,监控应用服务器内存泄露等问题: [JVM远程监控设置] 1.打开WAS控制台:https://ip:port/ibm/console/login.do 2.进入路径 ...
Android之build.prop属性详解
注:本篇文章是基于MSD648项目(AndroidTV)的prop进行说明. Android版本:4.4.4 内核版本:3.10.86 1.生成build.prop build.prop的生成是由ma ...
spring boot快速入门 10: 日志使用
第一步:pom 文件 <?xml version="1.0" encoding="UTF-8"?> <project xmlns=" ...
c++ 网络编程（十） LINUX/windows 异步通知I/O模型与重叠I/O模型附带示例代码
原文作者:aircraft 原文链接:https://www.cnblogs.com/DOMLX/p/9662931.html 一.异步IO模型(asynchronous IO) (1)什么是异步I/ ...
用table布局和div布局的区别
table布局的渲染是将整个table全部渲染出来,如果网路不给力的情况下,整个table会卡死在页面div布局的话,页面渲染,会一个一个的div渲染,网页出现会一个一个出来,不管网速怎样,不会全局卡 ...
构建流式应用—RxJS详解[转]
目录常规方式实现搜索功能 RxJS · 流 Stream RxJS 实现原理简析观察者模式迭代器模式 RxJS 的观察者 + 迭代器模式 RxJS 基础实现 Observable Observe ...
OpenTLD在VS2012和opencv246编译通过
最近看到了TLD的跟踪视频,觉得很有意思,刚好最近在看行人检测所以就打算下载源码玩一玩,因为源码是Linux版本的(原作者写的是C++和MATLAB的混合编程)C++源码可以在我的博客TLD(一种目标 ...
[转]a-mongodb-tutorial-using-c-and-asp-net-mvc
本文转自:http://www.joe-stevens.com/2011/10/02/a-mongodb-tutorial-using-c-and-asp-net-mvc/ In this post ...
springmvc 权限测试版
参考博文 https://blog.csdn.net/u011277123/article/details/68940939 1.Listener加载权限信息 2.interceptor验证权限测试 ...

Hadoop & Spark & Hive & HBase

HistoryServer

Hadoop & Spark & Hive & HBase的更多相关文章

随机推荐

热门专题