【发布时间】:2015-04-07 06:28:53
【问题描述】:
我刚刚复制了spark streaming wodcount python代码,使用spark-submit在Spark集群中运行wordcount python代码,但是显示如下错误:
py4j.protocol.Py4JJavaError: An error occurred while calling o23.loadClass.
: java.lang.ClassNotFoundException: org.apache.spark.streaming.kafka.KafkaUtilsPythonHelper
at java.net.URLClassLoader$1.run(URLClassLoader.java:366)
at java.net.URLClassLoader$1.run(URLClassLoader.java:355)
at java.security.AccessController.doPrivileged(Native Method)
at java.net.URLClassLoader.findClass(URLClassLoader.java:354)
我确实构建了 jar spark-streaming-kafka-assembly_2.10-1.4.0-SNAPSHOT.jar。我使用以下脚本提交: bin/spark-submit /data/spark-1.3.0-bin-hadoop2.4/wordcount.py --master spark://192.168.100.6:7077 --jars /data/spark-1.3.0-bin-hadoop2 .4/kafka-assembly/target/spark-streaming-kafka-assembly_*.jar。
提前致谢!
【问题讨论】:
标签: apache-spark apache-kafka spark-streaming spark-streaming-kafka