How it works?
Notebooks for Apache Spark completes the OVHcloud Data Processing product by leveraging our existing serverless Apache Spark service and bringing the capability to run on-demand Python Apache Spark jobs from your Jupyter Notebooks.
Data scientist will find a familiar data science experience through Jupyter Notebooks without the hassle of setting up the Apache Spark infrastructure.
How to opt-in this lab?
Nothing simpler! We have put together several guides to help setup and use your first notebook.
First, follow our getting started guide to learn how to create Notebooks for Apache Spark.
Then, try our data cleaning tutorial to hone your skills.
Features & benefits
Accelerate time to market
- As Data Scientists or Developers, benefit from Jupyter live code editor very simply
- No hassle of setting up Apache Spark infrastructure
- Launch your Jupyter notebook in minutes, and directly launch your Apache Spark jobs on demand
- Accelerate your data project time to deliver
Ease of use
- Easy to use Control Panel as well as a comprehensive API
About Notebooks for Apache Spark
- Alpha launched in April 2023
- DC availability: GRA
- Apache Spark 3.4.0
How to share your feedback?
Feel free to engage with us and the community on Discord.
You will find our #data-processing channel under the Data Analytics - Public Cloud section.
Limitations
This product being in an alpha testing phase comes with a number of limitations:
- Kernels are limited to the latest supported version of Spark (currently 3.4.0).
- Notebooks can only be created in "Public access".
- Environment are not backed up when stopping a notebook. Remember to save your files before stopping the Jupyter Lab.