presto mysql连接器_Apache Presto KAFKA连接器

最新推荐文章于 2024-09-26 04:00:34 发布

吴双无敌

最新推荐文章于 2024-09-26 04:00:34 发布

阅读量218

点赞数

文章标签： presto mysql连接器

本文链接：https://blog.csdn.net/weixin_35714577/article/details/114909493

版权

本文介绍了如何通过Presto的Kafka Connector访问Apache Kafka中的数据。首先，详细步骤包括下载并安装Apache ZooKeeper和Kafka，启动它们。接着，使用tpch-kafka工具预加载TPCH数据到Kafka主题。最后，配置Presto服务器以连接到Kafka，并在Presto CLI中查询加载的数据。

摘要由CSDN通过智能技术生成

Presto的Kafka Connector可以使用Presto访问Apache Kafka的数据。

学习提醒

下载并安装最新版本的以下Apache项目。

Apache ZooKeeper

Apache卡夫卡

启动ZooKeeper

使用以下命令启动ZooKeeper服务器。

$ bin/zookeeper-server-start.sh config/zookeeper.properties

现在，ZooKeeper在2181启动端口。

开始卡夫卡

使用以下命令在另一个终端中启动Kafka。

$ bin/kafka-server-start.sh config/server.properties

卡夫卡启动后，使用端口号9092。

TPCH数据

下载tpch-kafka

$ curl -o kafka-tpch

https://repo1.maven.org/maven2/de/softwareforge/kafka_tpch_0811/1.0/kafka_tpch_

0811-1.0.sh

现在您已经使用上述命令从Maven中心下载了加载程序。你会得到类似的回应如下。

% Total % Received % Xferd Average Speed Time Time Time Current

Dload Upload Total Spent Left Speed

0 0 0 0 0 0 0 0 --:--:-- 0:00:01 --:--:-- 0

5 21.6M 5 1279k 0 0 83898 0 0:04:30 0:00:15 0:04:15 129k

6 21.6M 6 1407k 0 0 86656 0 0:04:21 0:00:16 0:04:05 131k

24 21.6M 24 5439k 0 0 124k 0 0:02:57 0:00:43 0:02:14 175k

24 21.6M 24 5439k 0 0 124k 0 0:02:58 0:00:43 0:02:15 160k

25 21.6M 25 5736k 0 0 128k 0 0:02:52 0:00:44 0:02:08 181k

………………………..

然后，使用以下命令使其可执行，

$ chmod 755 kafka-tpch

运行tpch-kafka

运行kafka-tpch程序，使用以下命令预先加载一些具有tpch数据的主题。

查询

$ ./kafka-tpch load --brokers localhost:9092 --prefix tpch. --tpch-type tiny

结果

2016-07-13T16:15:52.083+0530 INFO main io.airlift.log.Logging Logging

to stderr

2016-07-13T16:15:52.124+0530 INFO main de.softwareforge.kafka.LoadCommand

Processing tables: [customer, orders, lineitem, part, partsupp, supplier,

nation, region]

2016-07-13T16:15:52.834+0530 INFO pool-1-thread-1

de.softwareforge.kafka.LoadCommand Loading table "customer" into topic "tpch.customer"...

2016-07-13T16:15:52.834+0530 INFO pool-1-thread-2

de.softwareforge.kafka.LoadCommand Loading table "orders" into topic "tpch.orders"...

2016-07-13T16:15:52.834+0530 INFO pool-1-thread-3

de.softwareforge.kafka.LoadCommand Loading table "lineitem" into topic "tpch.lineitem"...

2016-07-13T16:15:52.834+0530 INFO pool-1-thread-4

de.softwareforge.kafka.LoadCommand Loading table "part" into topic "tpch.part"...

………………………

……………………….

现在，Kafka表客户，订单，供应商等都使用tpch加载。

添加配置设置

我们在Presto服务器上添加以下Kafka连接器配置设置。

connector.name = kafka

kafka.nodes = localhost:9092

kafka.table-names = tpch.customer,tpch.orders,tpch.lineitem,tpch.part,tpch.partsupp,

tpch.supplier,tpch.nation,tpch.region

kafka.hide-internal-columns = false

在上述配置中，使用Kafka-tpch程序加载Kafka表。

启动Presto CLI

使用以下命令启动Presto CLI，

$ ./presto --server localhost:8080 --catalog kafka —schema tpch;

这里“tpch”是Kafka连接器的架构，您将收到以下回复。

presto:tpch>

列表

以下查询列出了“tpch”模式中的所有表。

查询

presto:tpch>show tables;

结果

Table

----------

customer

lineitem

nation

orders

part

partsupp

region

supplier

描述客户表

以下查询描述“客户”表。

查询

presto:tpch>describe customer;

结果

Column | Type | Comment

-------------------+---------+---------------------------------------------

_partition_id | bigint | Partition Id

_partition_offset | bigint | Offset for the message within the partition

_segment_start | bigint | Segment start offset

_segment_end | bigint | Segment end offset

_segment_count | bigint | Running message count per segment

_key | varchar | Key text

_key_corrupt | boolean | Key data is corrupt

_key_length | bigint | Total number of key bytes

_message | varchar | Message text

_message_corrupt | boolean | Message data is corrupt

_message_length | bigint | Total number of message bytes

吴双无敌

关注

0
点赞
踩
0

收藏

觉得还不错? 一键收藏
0
评论
复制链接

分享到 QQ

分享到新浪微博

扫一扫