Portfolio - ONIBIYO Joshua Toluse
Loading...
Profile Image

ONIBIYO Joshua Toluse

Data Engineer, Data Pipeline Developer, Data Infrastructure Architect, Big Data Engineer, Cloud Data Engineer, Data Quality Assurance Engineer, Real-Time Data Processing Engineer, DevOps for Data Engineering

About Me

I am a passionate and proactive Data Engineer with over three years of experience designing and managing scalable data pipelines and systems. My expertise lies in leveraging cutting-edge technologies to efficiently extract, transform, and load (ETL) data, enabling organizations to unlock actionable insights and drive data-informed decisions.

I thrive on solving real-world problems through data, whether it’s building end-to-end data pipelines, optimizing data storage, or enabling advanced analytics. My goal is to create scalable, efficient, and reliable data systems that empower businesses to thrive in a data-driven world.

I am always eager to take on new challenges and contribute to impactful projects that push the boundaries of what’s possible with data. If you’re looking for a dedicated data professional who combines technical expertise with a problem-solving mindset, let’s connect and explore how we can work together to drive innovation and create meaningful change.

Name: ONIBIYO Joshua Toluse
Birthday: 29 April
Degree: Bachelor
Experience: 5 Years
Phone: +234 813 526 1562
Email: toluseonibiyo@gmail.com
Address: 11 Gbajumo Crescent, Off Adeniran Ogunsanya, Surulere, Lagos, Nigeria
Freelance: Available

5

Years of

Experience

6

Happy

Clients

13

Completed

Projects

Skills

Python

72%

SQL

81%

Apache Spark

69%

GitHub

78%

Power BI

52%

Apache Airflow

66%

AWS

57%

Microsoft Azure

66%

Google Cloud Platform (GCP)

84%

Docker

65%

Experience

Data Engineer (Contract)

NYC Hospital | Nov. 2024 - Jan. 2025

Designed and implemented an ETL pipeline to process payroll, employee, agency, and title data, storing it in an AWS RDS PostgreSQL data warehouse. The pipeline extracted raw data from S3 buckets and transformed it using AWS Glue to address data quality issues such as null values, typecasting, and field standardization. Once cleaned and validated, the data was loaded into PostgreSQL tables for further analysis. To automate the process, I developed AWS Glue Notebook scripts for extraction, transformation, and loading, and created stored procedures to generate aggregate tables that answered key business questions.

Technologies: AWS Glue, AWS S3, AWS RDS (PostgreSQL), ETL Pipelines, Data Warehousing, Python (for scripting), SQL

Data Engineer (Contract)

Yanki eCommerce | Oct. 2024 - Nov. 2024

Designed and implemented a data pipeline to clean, transform, and load raw e-commerce data into a structured PostgreSQL database. Created normalized tables and enforced data integrity using primary and foreign keys. Automated data ingestion using Python, Pandas, and psycopg2, ensuring seamless data flow and storage. Delivered a scalable database solution, enabling efficient data analysis and decision-making for Yanki Ecommerce.

Technologies: Python, Pandas, PostgreSQL, ETL Pipelines, Database Design

Data Engineer (Contract)

GoalBet | Jun. 2024 - Jun. 2024

Designed and implemented an automated data pipeline to collect, clean, and load football match data (Premiership, Championship, and League 1) from an external API into a PostgreSQL database. Delivered a scalable and reusable solution for ongoing data collection, enabling GoalBet to analyze football match trends and make data-driven decisions.

Technologies Used: Python, Pandas, PostgreSQL, SQLAlchemy, REST APIs, Data Cleaning, ETL Pipelines

Services

Data Pipeline Development

Design and build scalable data pipelines using Python, SQL, and Apache Spark to process and transform large datasets efficiently.

Cloud Data Solutions

Implement cloud-based data solutions on AWS, Microsoft Azure, and GCP, leveraging services like S3, RDS, Glue, and BigQuery.

Data Orchestration with Airflow

Automate and orchestrate complex workflows using Apache Airflow to ensure seamless data processing and task scheduling.

Data Visualization with Power BI

Create interactive dashboards and reports using Power BI to provide actionable insights for business decision-making.

Big Data Processing with Spark

Process and analyze large-scale datasets using Apache Spark for real-time and batch processing.

Containerization with Docker

Containerize data applications and workflows using Docker for consistent deployment across environments.

Version Control with GitHub

Manage code repositories and collaborate on projects using GitHub for efficient version control and team collaboration.

Data Warehousing & SQL

Design and optimize data warehouses and write complex SQL queries for data extraction, transformation, and analysis.

Portfolio

  • All
  • On-Prem Solutions
  • Cloud Solutions
  • Hybrid Solutions
NYC Hospital ETL Pipeline

AWS Glue, PostgreSQL, PySpark, AWS S3, Boto3, PL/pgSQL

Yank eCommerce Data Pipeline

Python, Pandas, PostgreSQL

AgroFarm - Exchange Rate Data Extraction and Database Integration

API, Python, Pandas, PostgreSQL

GoalBet Football Data Pipeline

Python, REST APIs, PostgreSQL

Data Scraping and Snowflake Integration

Web Scraping, Snowflake DW, Python

AliExpress Data Scraping and Database Integration

Selenium, BeautifulSoup, PostgreSQL, SQLAlchemy, Python

Divy Trips - Docker and Airflow ETL Pipeline for ClickHouse

Apache Airflow, ClickHouse, PostgreSQL, Docker, Python, SQLAlchemy

Zulu Bank Data Storage Optimization

Python, Pandas, PostgreSQL

Contact Me

I'd love to hear from you! Whether you have a question, a project idea, or just want to connect, feel free to reach out. I'll get back to you as soon as possible.

© Joshua Onibiyo. All Rights Reserved.