Logo Questions Linux Laravel Mysql Ubuntu Git Menu
 

New posts in apache-spark

Spark error: "ERROR Utils: Exception while deleting Spark temp dir:"

How do I write HDF5 files from Apache Spark?

apache-spark hdf5

Do I need to use "mergeSchema" option in spark with parquet if I am passing in a schema explicitly?

apache-spark parquet

Adding a Column to Spark Table via SQL ALTER TABLE command

How to use long user ID in PySpark ALS

Query spark on JSON object stored on Cassandra DB

difference between spark.executor.memoryOverhead and spark.memory.offHeap.size

apache-spark memory jvm

spark-shell - How to avoid suppressing of elided stack trace (Exceptions)

BSON structure created by Apache Spark and MongoDB Hadoop-Connector

How to read a csv file from s3 bucket using pyspark

access objects in pyspark user-defined function from outer scope, avoid PicklingError: Could not serialize object

Does spark read the same file twice, if two stages are using the same DataFrame?

PhoenixOutputFormat not found when running a Spark Job on CDH 5.4 with Phoenix 4.5

multiple contact points in the spark cassandra connector

cassandra apache-spark

Bluemix spark-submit -- How to secure credentials needed by my Scala jar

Vegas (Scala/Spark/Vega) color every data point

Must include log4J, but it is causing errors in Apache Spark shell. How to avoid errors?