快捷導(dǎo)航

基于SpringBoot?使用?Flink?收發(fā)Kafka消息的示例詳解

更新時(shí)間：2023年01月07日 09:12:25 作者：dk168

這篇文章主要介紹了基于SpringBoot?使用?Flink?收發(fā)Kafka消息,本文通過(guò)示例代碼給大家介紹的非常詳細(xì)，對(duì)大家的學(xué)習(xí)或工作具有一定的參考借鑒價(jià)值，需要的朋友可以參考下

前言

這周學(xué)習(xí)下Flink相關(guān)的知識(shí)，學(xué)習(xí)到一個(gè)讀寫Kafka消息的示例，自己動(dòng)手實(shí)踐了一下，別人示例使用的是普通的Java Main方法，沒有用到spring boot. 我們?cè)趯?shí)際工作中會(huì)使用spring boot。因此我做了些加強(qiáng)，把流程打通了，過(guò)程記錄下來(lái)。

準(zhǔn)備工作

首先我們通過(guò)docker安裝一個(gè)kafka服務(wù)，參照Kafka的官方知道文檔
https://developer.confluent.io/tutorials/kafka-console-consumer-producer-basics/kafka.html主要的是有個(gè)docker-compose.yml文件

---
version: '2'

services:
  zookeeper:
    image: confluentinc/cp-zookeeper:7.3.0
    hostname: zookeeper
    container_name: zookeeper
    ports:
      - "2181:2181"
    environment:
      ZOOKEEPER_CLIENT_PORT: 2181
      ZOOKEEPER_TICK_TIME: 2000

  broker:
    image: confluentinc/cp-kafka:7.3.0
    hostname: broker
    container_name: broker
    depends_on:
      - zookeeper
    ports:
      - "29092:29092"
    environment:
      KAFKA_BROKER_ID: 1
      KAFKA_ZOOKEEPER_CONNECT: 'zookeeper:2181'
      KAFKA_LISTENER_SECURITY_PROTOCOL_MAP: PLAINTEXT:PLAINTEXT,PLAINTEXT_HOST:PLAINTEXT
      KAFKA_ADVERTISED_LISTENERS: PLAINTEXT://broker:9092,PLAINTEXT_HOST://localhost:29092
      KAFKA_OFFSETS_TOPIC_REPLICATION_FACTOR: 1
      KAFKA_GROUP_INITIAL_REBALANCE_DELAY_MS: 0

docker compose up -d
就可以把kafka docker 環(huán)境搭起來(lái)，
使用以下命令，創(chuàng)建一個(gè)flink.kafka.streaming.source的topic
docker exec -t broker kafka-topics --create --topic flink.kafka.streaming.source --bootstrap-server broker:9092
然后使用命令，就可以進(jìn)入到kafka機(jī)器的命令行
docker exec -it broker bash
官方文檔示例中沒有-it, 運(yùn)行后沒有進(jìn)入broker的命令行，加上來(lái)才可以。這里說(shuō)明下

Flink我們打算直接采用開發(fā)工具運(yùn)行，暫時(shí)未搭環(huán)境，以體驗(yàn)為主。

開發(fā)階段

首先需要引入的包POM文件

    <properties>
        <jdk.version>1.8</jdk.version>
        <maven.compiler.source>8</maven.compiler.source>
        <maven.compiler.target>8</maven.compiler.target>
        <project.build.sourceEncoding>UTF-8</project.build.sourceEncoding>
        <spring-boot.version>2.7.7</spring-boot.version>
        <flink.version>1.16.0</flink.version>
    </properties>
    <dependencyManagement>
        <dependencies>
            <dependency>
                <groupId>org.springframework.boot</groupId>
                <artifactId>spring-boot-dependencies</artifactId>
                <version>${spring-boot.version}</version>
                <type>pom</type>
                <scope>import</scope>
            </dependency>
        </dependencies>
    </dependencyManagement>
    <dependencies>
        <dependency>
            <groupId>org.springframework.boot</groupId>
            <artifactId>spring-boot-starter</artifactId>
        </dependency>
        <dependency>
            <groupId>org.projectlombok</groupId>
            <artifactId>lombok</artifactId>
            <optional>true</optional>
        </dependency>
        <dependency>
            <groupId>org.apache.flink</groupId>
            <artifactId>flink-java</artifactId>
            <version>${flink.version}</version>
            <scope>provided</scope>
        </dependency>
        <dependency>
            <groupId>org.apache.flink</groupId>
            <artifactId>flink-clients</artifactId>
            <version>${flink.version}</version>
            <scope>provided</scope>
        </dependency>
        <dependency>
            <groupId>org.apache.flink</groupId>
            <artifactId>flink-streaming-java</artifactId>
            <version>${flink.version}</version>
            <scope>provided</scope>
        </dependency>
        <dependency>
            <groupId>org.apache.flink</groupId>
            <artifactId>flink-connector-kafka</artifactId>
            <version>${flink.version}</version>
            <scope>provided</scope>
        </dependency>
    </dependencies>

這里我們使用Java8，本來(lái)想使用Spring Boot 3的，但是Spring Boot 3 最低需要Java17了，目前Flink支持Java8和Java11，所以我們使用Spring Boot 2, Java 8來(lái)開發(fā)。

spring-boot-starter 我們就一個(gè)命令行程序，所以用這個(gè)就夠了
lombok 用來(lái)定義model
flink-java, flink-clients, flink-streaming-java 是使用基本組件，缺少flink-clients編譯階段不會(huì)報(bào)錯(cuò)，運(yùn)行的時(shí)候會(huì)報(bào)java.lang.IllegalStateException: No ExecutorFactory found to execute the application.
flink-connector-kafka 是連接kafka用
我們這里把provided，打包的時(shí)候不用打包flink相關(guān)組件，由運(yùn)行環(huán)境提供。但是IDEA運(yùn)行的時(shí)候會(huì)報(bào)java.lang.NoClassDefFoundError: org/apache/flink/streaming/util/serialization/DeserializationSchema，
在運(yùn)行的configuration上面勾選上“add dependencies with provided scope to classpath”可以解決這個(gè)問(wèn)題。

主要代碼

@Component
@Slf4j
public class KafkaRunner implements ApplicationRunner
{
    @Override
    public void run(ApplicationArguments args) throws Exception {
        try{

            /****************************************************************************
             *                 Setup Flink environment.
             ****************************************************************************/

            // Set up the streaming execution environment
            final StreamExecutionEnvironment streamEnv
                    = StreamExecutionEnvironment.getExecutionEnvironment();

            /****************************************************************************
             *                  Read Kafka Topic Stream into a DataStream.
             ****************************************************************************/

            //Set connection properties to Kafka Cluster
            Properties properties = new Properties();
            properties.setProperty("bootstrap.servers", "localhost:29092");
            properties.setProperty("group.id", "flink.learn.realtime");

            //Setup a Kafka Consumer on Flnk
            FlinkKafkaConsumer<String> kafkaConsumer =
                    new FlinkKafkaConsumer<>
                            ("flink.kafka.streaming.source", //topic
                                    new SimpleStringSchema(), //Schema for data
                                    properties); //connection properties

            //Setup to receive only new messages
            kafkaConsumer.setStartFromLatest();

            //Create the data stream
            DataStream<String> auditTrailStr = streamEnv
                    .addSource(kafkaConsumer);

            //Convert each record to an Object
            DataStream<Tuple2<String, Integer>> userCounts
                    = auditTrailStr
                    .map(new MapFunction<String,Tuple2<String,Integer>>() {

                        @Override
                        public Tuple2<String,Integer> map(String auditStr) {
                            System.out.println("--- Received Record : " + auditStr);
                            AuditTrail at = new AuditTrail(auditStr);
                            return new Tuple2<String,Integer>(at.getUser(),at.getDuration());
                        }
                    })

                    .keyBy(0)  //By user name
                    .reduce((x,y) -> new Tuple2<String,Integer>( x.f0, x.f1 + y.f1));

            //Print User and Durations.
            userCounts.print();

            /****************************************************************************
             *                  Setup data source and execute the Flink pipeline
             ****************************************************************************/
            //Start the Kafka Stream generator on a separate thread
            System.out.println("Starting Kafka Data Generator...");
            Thread kafkaThread = new Thread(new KafkaStreamDataGenerator());
            kafkaThread.start();

            // execute the streaming pipeline
            streamEnv.execute("Flink Windowing Example");

        }
        catch(Exception e) {
            e.printStackTrace();
        }
    }
}

簡(jiǎn)單說(shuō)明下程序
DataStream auditTrailStr = streamEnv
.addSource(kafkaConsumer);
就是接通了Kafka Source

        Thread kafkaThread = new Thread(new KafkaStreamDataGenerator());
        kafkaThread.start();

這段代碼是另外開一個(gè)線程往kafka里面去發(fā)送文本消息
我們?cè)谶@個(gè)示例中就是一個(gè)線程發(fā)，然后flink就讀出來(lái)，然后統(tǒng)計(jì)出每個(gè)用戶的操作時(shí)間。
auditTrailStr.map 就是來(lái)進(jìn)行統(tǒng)計(jì)操作。

運(yùn)行效果

可以看到Kafka一邊發(fā)送，然后我們就一邊讀出來(lái)，然后就統(tǒng)計(jì)出了每個(gè)用戶的時(shí)間。

總結(jié)

本文只是簡(jiǎn)單的打通了幾個(gè)環(huán)節(jié)，對(duì)于flink的知識(shí)沒有涉及太多，算是一個(gè)環(huán)境入門。后面學(xué)習(xí)更多的以后我們?cè)偕钊胄﹣?lái)記錄flink. 示例代碼會(huì)放到 https://github.com/dengkun39/redisdemo.git spring-boot-flink 文件夾。

您可能感興趣的文章: