项目01-flume、kafka与hdfs日志流转
2024-08-31 07:11:17
项目01-flume、kafka与hdfs日志流转
1、启动kafka集群
$>xkafka.sh start
3、创建kafka主题
kafka-topics.sh --zookeeper s102:2181
--create
--topic topic-umeng-raw-logs2
--replication-factor 3
--partitions 4
注意:kafka主题不要使用“_”,可以使用“-”。
4、配置flume,收集日志到kafka
在nginx web服务器节点(这里是s101和s102)上安装flume软件,编写配置文件。
在/soft/flume/conf下创建umeng_nginx_to_kafka.conf文件,内容如下:
a1.sources = r1
a1.channels = c1
a1.sinks = k1
a1.sources.r1.type = exec
a1.sources.r1.command = tail -F /usr/local/openresty/nginx/logs/access.log
a1.channels.c1.type = memory
a1.channels.c1.capacity = 10000
a1.sinks.k1.type = org.apache.flume.sink.kafka.KafkaSink
a1.sinks.k1.kafka.topic = topic-umeng-raw-logs2
a1.sinks.k1.kafka.bootstrap.servers = s102:9092
a1.sinks.k1.kafka.flumeBatchSize = 20
a1.sinks.k1.kafka.producer.acks = 1
a1.sinks.k1.kafka.producer.linger.ms = 0
a1.sources.r1.channels=c1
a1.sinks.k1.channel=c1
5、启动flume进程
$>flume-ng agent -f /soft/flume/conf/umeng_nginx_to_kafka.conf -n a1
6、启动kafka控制条消费者,查看是否能够接收到日志
$>kafka-console-consumer.sh --zookeeper s102:2181 --topic topic-umeng-raw-logs2
7、配置flume,收集kafka消息到hdfs
创建/soft/flume/conf/umeng-kakfa-to-hdfs.conf文件,内容如下:
a1.sources = r1
a1.channels = c1
a1.sinks = k1
a1.sources.r1.type = org.apache.flume.source.kafka.KafkaSource
a1.sources.r1.batchSize = 5000
a1.sources.r1.batchDurationMillis = 2000
a1.sources.r1.kafka.bootstrap.servers = s102:9092
a1.sources.r1.kafka.topics = topic-umeng-raw-logs2
a1.sources.r1.kafka.consumer.group.id = g10
a1.channels.c1.type=memory
a1.sinks.k1.type = hdfs
a1.sinks.k1.hdfs.path = /user/centos/umeng_big11/raw-logs/%Y%m/%d/%H%M
a1.sinks.k1.hdfs.filePrefix = events-
#round控制目录
a1.sinks.k1.hdfs.round = true
a1.sinks.k1.hdfs.roundValue = 1
a1.sinks.k1.hdfs.roundUnit = minute
#控制文件
a1.sinks.k1.hdfs.rollInterval = 30
a1.sinks.k1.hdfs.rollSize = 10240
a1.sinks.k1.hdfs.rollCount = 500
a1.sinks.k1.hdfs.useLocalTimeStamp = true
a1.sinks.k1.hdfs.fileType = DataStream
a1.sources.r1.channels=c1
a1.sinks.k1.channel=c1
8、启动flume进程,收集kafka消息到hdfs
8.1 启动hdfs集群
$>start-dfs.sh
8.2 启动flume,指定收集文件
$>flume-ng agent -f /soft/flume/conf/umeng-kafka-to-hdfs.conf -n a1
9、启动手机端程序发送日志,观察kafka是否接收到
最新文章
- 如果你想深刻理解ASP.NET Core请求处理管道,可以试着写一个自定义的Server
- CodeForces 676D代码 哪里有问题呢?
- Linux上从Java程序中调用C函数
- XMPP环境搭建
- Java命令提示符编译
- function foo(){}、(function(){})、(function(){}())等函数区别分析
- 使用eclipse编译调试c++
- WPF passwordbox 圆角制作
- html a标签 图片边框和点击后虚线框的有关问题
- React组件的生命周期各环节运作流程
- 使用Lock实现信号量
- 杭电oj find your present (2)
- 参加persist.sys物业写权限的方法
- pkgmgmt: Comparison between different Linux Systems..
- Delphi中建立指定大小字体和读取该字体点阵信息的函数(转)
- angr进阶(6)绕过反调试
- Chrome中的哪些端口是限制使用的?
- Lanczos Algorithm and it's Parallelization Stragegy
- window.open和window.showModalDialog
- Virut样本取证特征
热门文章
- ERROR (UnicodeEncodeError): 'ascii' codec can't encode character u'\uff08' in position 9: ordinal not in range(128)
- freemarker 定义公共header
- php __CLASS__、get_class()与get_called_class()的区别
- saltstack一键部署高可用
- my.时空_物价
- Android文件/文件夹选择器(支持多选操作),已封装为lib库,直接添加依赖即可。
- ";i++";和“++i”的区别
- poi 详细demo
- Unity [SerializeField]
- 性能测试工具LoadRunner08-LR之Virtual User Generator 检查点