Skip to content
View ingridevv's full-sized avatar
🧬
CRISPR-Cas9
🧬
CRISPR-Cas9

Block or report ingridevv

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ingridevv/README.md

Ingrid da Silva

Data Engineer | Google Cloud Certified Professional Data Engineer

Google Cloud Professional Data Engineer

I design and build cloud-native data platforms that power reliable analytical systems. My work spans scalable data infrastructure, distributed data processing, modern ELT architectures, and analytics engineering on Google Cloud. I am interested in advancing data engineering and AI for genomics, bioinformatics, and scientific computing.

Expertise

Architecture Cloud-native Data Platforms Distributed Data Processing Scalable Data Infrastructure Lakehouse Architecture Big Data

Data Engineering ETL/ELT Data Pipelines Batch Processing Analytics Engineering Data Modeling Data Warehousing

Cloud & AI Google Cloud BigQuery Vertex AI Generative AI

Programming Python SQL

Infrastructure Apache Airflow Apache Spark Docker Terraform CI/CD

Featured Projects

End-to-end genomic data platform for processing and analyzing more than 1.1 million GWAS variants through reproducible cloud-native data pipelines and infrastructure automation.

Technologies: dbtSnowflakePythonApache AirflowDockerTerraformGitHub Actions


Certifications

  • Google Cloud Certified Professional Data Engineer
  • AWS Developer – Associate
  • Oracle AI Foundations Associate
  • AWS Certified Cloud Practitioner

Connect

LinkedIn


Pinned Loading

  1. genomic-disease-risk-pipeline genomic-disease-risk-pipeline Public

    An end-to-end data engineering solution for genomic discovery. Using a multi-tiered Medallion approach, the pipeline orchestrates the transformation of 1.1M+ genetic variants into actionable biolog…

    Python

  2. data-engineering-zoomcamp data-engineering-zoomcamp Public

    Data Engineering projects developed during the DataTalks.Club Zoomcamp 2026.

    Jupyter Notebook 1