#102 · Primary category: Education & Research

data-science-ipython-notebooks

aws big-data caffe data-science deep-learning hadoop kaggle keras machine-learning mapreduce matplotlib numpy pandas python scikit-learn scipy spark tensorflow theano

Data science Python notebooks: Deep learning (TensorFlow, Theano, Caffe, Keras), scikit-learn, Kaggle, big data (Spark, Hadoop MapReduce, HDFS), matplotlib, pandas, NumPy, SciPy, Python essentials, AWS, and various command lines.

Project last updated:03/20/24

GitHub Stars

29.3K

Forks

8.0K

Contributors

13

License

Other

Why we included this project

These Jupyter notebooks work through the Python data science stack in a sensible order, starting with NumPy, pandas, and matplotlib, then moving into scikit-learn, deep learning with TensorFlow and Keras, and big-data tooling like Spark and Hadoop. It is a learning resource rather than a deployable application, so its value is in the worked examples: each notebook shows real code you can read before writing your own. The notebooks are grouped by topic, which makes it easy to jump straight to the library you are currently learning, and the Kaggle and business-analysis sections show how the pieces fit together on actual problems. New data scientists can use it to get up to speed on the fundamentals, and experienced practitioners will find it a quick way to recall syntax across many tools without digging through separate docs.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category