gedeck/practical-statistics-for-data-scientists

Code repository for O'Reilly book

View on GitHub ↗Jump to charts ↓

Summary Information

Updated 25 minutes ago
Added to GitGenius on September 20th, 2026
Created on February 17th, 2020
Open Issues & Pull Requests: 4 (+0)
GitHub issues: Enabled
Number of forks: 1,992
Total Stargazers: 3,382 (+0)
Total Subscribers: 75 (+0)

Repository Insights (GitGenius)

Median issue/PR response: 3.9 hours
Mean response time: 3.9 hours
90th percentile: 3.9 hours
Tracked items: 1

Most active contributors

Sign in to see contributor activity.

Related repositories by overlapping contributors

No overlapping-contributor repos identified yet.

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Open issues: 1
New in 7 days: 0
Closed in 7 days: 0
Avg open age: 87 days
Stale 30+ days: 1
Stale 90+ days: 0

Recent activity

Opened in 7 days: 0
Closed in 7 days: 0
Comments in 7 days: 0
Events in 7 days: 0

Top labels

No label distribution available yet.

Most active issues this week

No issue events were indexed in the last 7 days.

Detailed Description

Practical Statistics for Data Scientists is a code repository accompanying an O'Reilly book that provides implementations and examples for statistical concepts and techniques relevant to data science work.

The repository addresses the need for practitioners to understand and apply statistical methods in real-world data science projects. It works by pairing theoretical statistical concepts with executable code examples, allowing readers to see how these methods translate into practice. The examples are organized to correspond with chapters in the accompanying book, making it straightforward to follow along with the material and experiment with different approaches.

This repository suits data scientists and analysts who want to deepen their understanding of statistics beyond theoretical knowledge, particularly those working through the O'Reilly book or seeking practical implementations of statistical techniques. It works best for self-directed learning and as a reference when applying statistical methods to actual projects. The code-first approach makes it valuable for those who learn better by reading and modifying working examples rather than studying formulas alone.

The project maintains a stable collection of examples with periodic updates to keep dependencies current and fix issues as they arise. Contributions from the community are accepted and integrated into the codebase. The repository serves primarily as a reference implementation rather than an actively developed software library, with changes focused on ensuring the examples remain functional and relevant rather than expanding the scope of covered topics.