Skip to content
View cedekpoole's full-sized avatar
🤯
🤯

Block or report cedekpoole

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
cedekpoole/README.md

Hi, I'm Cameron Poole | MSc Data Science student | SQL · Python · data analysis · modelling 📊

I'm completing an MSc Data Science at Kingston University after a First-Class BA in Philosophy from Newcastle University.

I primarily work with SQL and Python, and I'm interested in applied data work: cleaning data, writing queries, building small pipelines, checking assumptions and using analysis or modelling to answer meaningful questions.

My current projects cover data pipelines, dbt models, statistical modelling, data quality checks and public datasets. I also have front-end software development experience, including React and scientific data visualisation work.

I'm looking for entry-level data roles in London across applied analytics, data analysis, data quality, public sector data, risk/evaluation, analytics engineering and junior data engineering.

MSc Data Science modules

  • Applied Data Programming — 81/100
  • Data Analytics and Visualisation — 99/100
  • Machine Learning and Deep Learning — 82/100
  • Databases and Data Management — 96/100
  • Dissertation — in progress

LinkedIn

name: Cameron Poole
location: London, UK
education: "MSc Data Science (Kingston, 2026) | BA Philosophy (1st, Newcastle)"

focus:
  - data analysis
  - applied analytics
  - statistical modelling
  - data quality
  - data pipelines

core_stack:
  - SQL
  - Python
  - R
  - Git

data_tools:
  - pandas
  - NumPy
  - scikit-learn
  - dbt
  - BigQuery
  - Azure Blob Storage
  - Snowflake
  - PostgreSQL
  - Oracle SQL

modelling_and_analysis:
  - regression
  - logistic regression
  - classification metrics
  - model evaluation
  - data cleaning
  - validation checks

visualisation_and_dev:
  - Tableau
  - Looker Studio
  - matplotlib
  - seaborn
  - JavaScript
  - TypeScript
  - React

seeking: "Entry-level data roles — London, hybrid preferred"

Current project

Building a housing and commuting analytics project using public property sales, EPC and rail data. The project asks how much more home an extra 10 minutes on the train buys a first-time buyer, focusing on the Great Eastern corridor into London Liverpool Street. It will include a tested Python pipeline, Snowflake and dbt models, data quality and record matching checks, and statistical modelling of the relationship between floor space and commute time.

Recently completed

Built a TfL transport analytics mart using Python, Azure Blob Storage, BigQuery, dbt and GitHub Actions. The project ingests Tube status data from the TfL API, stores raw JSON snapshots, loads the data into BigQuery and builds tested dbt models for line status and reliability analysis.

Other relevant work

  • Environment Agency dissertation on compliance history and serious-breach risk scoring.
  • Statistical modelling using regression, logistic regression, diagnostics and classification metrics.
  • Rail punctuality warehouse using Oracle SQL, star schema design and Tableau reporting.
  • Scientific visualisation work with React and Highcharts, including PCA and RNA-seq volcano plot components.

📫 How to reach me


Pinned Loading

  1. portfolio-website portfolio-website Public

    Personal Portfolio 💫

    TypeScript

  2. statistical-modelling-r statistical-modelling-r Public

    Statistical modelling project in R using regression, diagnostics and classification evaluation.

    R

  3. rna-seq-plots rna-seq-plots Public

    An npm package containing rna-seq plots

    JavaScript

  4. tfl-analytics-mart tfl-analytics-mart Public

    End-to-end data pipeline for cleaning and modelling Tfl data in BigQuery with dbt.

    Python

  5. support2-quality-of-life-analysis support2-quality-of-life-analysis Public

    Python analysis of clinical data exploring quality-of-life outcomes, feature relationships and statistical patterns.

    Jupyter Notebook

  6. philosophy-writing philosophy-writing Public

    Selected philosophical works on biopolitics, psychoanalysis, conspiracy theory, uncertainty and systems of belief.