Our Community is getting an upgrade! To get everything ready for the relaunch, we’ll be placing the site in read-only mode starting September 21st. We really appreciate your understanding while we get things set up behind the scenes. Catch up on all the exciting details about the move here. Need help or have questions? Drop us a line at [email protected]
Created on 06-28-201708:58 PM - edited on 04-21-202606:29 AM by GrazittiAPI
Setting Up a Data Science Platform on HDP using Anaconda
Building a Data Science Platform using Anaconda needs to be
able to
Launch PySpark jobs on the cluster
Synchronize python libraries from vetted public
repositories
Isolate environments with specific dependencies
to run production jobs using an older version of a package whilst simultaneously
running new version of the package
Launching notebooks and PySpark jobs using
different kernels such as Python_2.7, Python_3.x, R, Scala