Flink水位线学习

该代码示例展示了如何在Flink中使用水位线策略处理顺序增长的数据,特别是在事件时间语义下。通过TumblingEventTimeWindows定义10秒事件时间窗口,并使用自定义的TimestampAssigner设定水位线,以确保数据的正确处理。
摘要由CSDN通过智能技术生成

Flink的水位线(顺序增长的数据)

package com.apple.flink.watermark;

import com.apple.flink.func.WaterSensorMapFunction;
import com.apple.flink.model.WaterSensor;
import org.apache.commons.lang3.time.DateFormatUtils;
import org.apache.flink.api.common.eventtime.SerializableTimestampAssigner;
import org.apache.flink.api.common.eventtime.WatermarkStrategy;
import org.apache.flink.streaming.api.datastream.KeyedStream;
import org.apache.flink.streaming.api.datastream.SingleOutputStreamOperator;
import org.apache.flink.streaming.api.datastream.WindowedStream;
import org.apache.flink.streaming.api.environment.StreamExecutionEnvironment;
import org.apache.flink.streaming.api.functions.windowing.ProcessWindowFunction;
import org.apache.flink.streaming.api.windowing.assigners.SlidingProcessingTimeWindows;
import org.apache.flink.streaming.api.windowing.assigners.TumblingEventTimeWindows;
import org.apache.flink.streaming.api.windowing.assigners.TumblingProcessingTimeWindows;
import org.apache.flink.streaming.api.windowing.time.Time;
import org.apache.flink.streaming.api.windowing.windows.TimeWindow;
import org.apache.flink.util.Collector;

/**
 * 水位线基于事件语义
 *  升序的水位线
 */
@SuppressWarnings("all")
public class WaterMarkDemo {
    public static void main(String[] args) throws Exception {
        StreamExecutionEnvironment environment = StreamExecutionEnvironment.getExecutionEnvironment();
        environment.setParallelism(1);
        SingleOutputStreamOperator<WaterSensor> sensorDS = environment.socketTextStream("hadoop102", 7777)
                .map(new WaterSensorMapFunction())
                .assignTimestampsAndWatermarks(WatermarkStrategy.<WaterSensor>forMonotonousTimestamps()
                        .withTimestampAssigner(new SerializableTimestampAssigner<WaterSensor>() {
                            @Override
                            public long extractTimestamp(WaterSensor waterSensor, long l) {
                                System.out.println("数据=" + waterSensor + ",recordTS--->" + l);
                                return waterSensor.getVc() * 1000L;
                            }
                        }));

        WindowedStream<WaterSensor, String, TimeWindow> sensorWS =
                sensorDS.keyBy(WaterSensor::getId)
     //指定事件时间语义的窗口
                .window(TumblingEventTimeWindows.of(Time.seconds(10)));//事件事件 窗口长度为10秒

        SingleOutputStreamOperator<String> processed = sensorWS.process(new ProcessWindowFunction<WaterSensor, String, String, TimeWindow>() {
            @Override
            public void process(String value, ProcessWindowFunction<WaterSensor, String, String, TimeWindow>.Context context, Iterable<WaterSensor> iterable, Collector<String> collector) throws Exception {
                long start = context.window().getStart();
                long end = context.window().getEnd();
                String startFormat = DateFormatUtils.format(start, "yyyy-MM-dd HH:mm:ss.SSS");
                String endFormat = DateFormatUtils.format(end, "yyyy-MM-dd HH:mm:ss.SSS");
                long count = iterable.spliterator().estimateSize();
                collector.collect("key=" + value + "窗口【" + startFormat + "," + endFormat + ")包含" + count + "条数据--->" + iterable.toString());
            }
        });
        processed.print("---->");
        environment.execute();
    }
}

默认是时间语义
升序的时间戳指定为WaterMark
评论
添加红包

请填写红包祝福语或标题

红包个数最小为10个

红包金额最低5元

当前余额3.43前往充值 >
需支付:10.00
成就一亿技术人!
领取后你会自动成为博主和红包主的粉丝 规则
hope_wisdom
发出的红包
实付
使用余额支付
点击重新获取
扫码支付
钱包余额 0

抵扣说明:

1.余额是钱包充值的虚拟货币,按照1:1的比例进行支付金额的抵扣。
2.余额无法直接购买下载,可以购买VIP、付费专栏及课程。

余额充值