接一下以一个示例配置来介绍一下如何以Flink连接HDFS

1. 依赖HDFS

pom.xml 添加依赖

    <dependency>
        <groupId>org.apache.flink</groupId>
        <artifactId>flink-hadoop-compatibility_2.11</artifactId>
        <version>${flink.version}</version>
    </dependency>
    <dependency>
        <groupId>org.apache.hadoop</groupId>
        <artifactId>hadoop-client</artifactId>
        <version>${hadoop.version}</version>
    </dependency>

2. 配置 HDFS

hdfs-site.xmlcore-site.xml放入到src/main/resources目录下面

3. 读取HDFS上面文件

  final ExecutionEnvironment env = ExecutionEnvironment.getExecutionEnvironment();
        DataSource<String> text = env.readTextFile("hdfs://flinkhadoop:9000/user/wuhulala/input/core-site.xml");

TIP

  1. 请关闭HDFS 权限,不关闭需要把认证copy到resources目录下
 <property>
        <name>dfs.permissions</name>
        <value>false</value>
    </property>
 

相关文章:

  • 2022-01-03
  • 2022-01-03
  • 2022-02-12
  • 2021-11-23
  • 2021-10-24
  • 2021-06-06
  • 2022-01-01
  • 2021-06-03
猜你喜欢
  • 2021-09-19
  • 2021-08-20
  • 2021-08-25
  • 2022-01-04
  • 2021-12-14
  • 2022-02-03
  • 2022-01-21
相关资源
相似解决方案