Skip to content
OnData.blog

OnData.blog

Menu

  • Articles
  • By topic
  • About
  • Linkedin
  • Facebook
  • twitter
  • RSS

#apachespark

The Data Lake at Sopra Steria

Following on my previous post, we have spent some time on building an internal Data Lake at Sopra Steria. The infrastructure is functional now and admitting its first users. Much has been said on building successful Data Science teams. Multidisciplinary

Pawel Plaszczak February 17, 2020February 18, 2020 Articles 4 Comments Read more

Recent Posts

  • Moving On April 4, 2026
  • Data Literacy: Six examples of bad data interpretation April 29, 2024
  • Porting PyTorch neural network to Amazon AWS June 30, 2022
  • Porting pyTorch cloud detection model to Amazon AWS S3 June 17, 2022
  • pushing data to AWS. SageMaker sucks. So does Anaconda June 14, 2022
  • Linear Regression: Killer App with 19-century maths January 19, 2022
  • Democratization of statistics: Chi2 for non-experts January 12, 2022
  • An approach to categorize multi-lingual phrases December 15, 2021
  • The implications of Scikit-learn bug #21455 November 29, 2021
  • Your model may be inaccurate November 25, 2021

Recent Posts

  • Moving On
  • Data Literacy: Six examples of bad data interpretation
  • Porting PyTorch neural network to Amazon AWS
  • Porting pyTorch cloud detection model to Amazon AWS S3
  • pushing data to AWS. SageMaker sucks. So does Anaconda

Recent Comments

  • Pawel Plaszczak on How to isolate data that constitutes a spike in histogram?
  • robert on How to isolate data that constitutes a spike in histogram?
  • Marcello Anselmi Tamburini on Your model may be inaccurate
  • C on Product Owner vs Product Manager vs Architect
  • Houcem on Don’t trust Data Science. Ask the people

Archives

  • April 2026
  • April 2024
  • June 2022
  • January 2022
  • December 2021
  • November 2021
  • October 2021
  • June 2021
  • April 2021
  • March 2021
  • February 2021
  • October 2020
  • July 2020
  • June 2020
  • April 2020
  • March 2020
  • February 2020
  • November 2019
  • October 2019
  • May 2019
  • April 2019
  • March 2019
  • February 2019
  • January 2019
  • December 2018
  • November 2018
  • October 2018
  • September 2018
  • July 2018
  • November 2016

Categories

  • Articles
  • General Public
  • Uncategorized

Meta

  • Log in
  • Entries feed
  • Comments feed
  • WordPress.org
Copyright © 2026 OnData.blog. All rights reserved. Theme Spacious by ThemeGrill. Powered by: WordPress.