Data Engineer · Backend Developer

ALEJANDRO

I build data architecture and the backend systems behind it — ETL pipelines, robust scrapers, and automated flows that turn messy, heterogeneous sources into usable information. This page is painted live by a GPU fluid simulation; move your cursor and leave a trace.

SCROLL ↓
About

Hi, I'm Alejandro — a Systems Engineering student (about to graduate) based in Medellín, Colombia, with hands-on experience in data architecture and backend development. I specialize in designing ETL processes, optimizing SQL Server queries, and automating the flow of information across complex, heterogeneous sources.

At Dapper I build scrapers and ingestion pipelines that track legal and regulatory data from government portals worldwide. Before that I worked full stack at TDP Solutions and as a Systems & Information Analyst at Universidad Nacional de Colombia. I'm comfortable across the stack and enjoy combining modern MLOps and Big Data practices — from PySpark and Airflow to self-hosted infrastructure with Docker.

Projects

Legal Data Scrapers — Dapper

As Data Engineer at Dapper, I design and maintain specialized scrapers that track new laws and regulations across government portals in multiple countries, plus ingestion pipelines that process and structure large volumes of legal documents (laws, decrees, resolutions) from heterogeneous sources. Built robust scraping strategies that absorb constant changes in government site structures to keep the data flowing.

WEB SCRAPINGDATA PIPELINESPYTHON

Corona — Payment Reconciliation

At TDP Solutions, I designed and implemented data-transformation logic in SQL Server through complex stored procedures, achieving automated unification and reconciliation of large-scale enterprise payments.

SQL SERVERSTORED PROCEDURESETL

Serlogistica — Invoice Intelligence

At TDP Solutions, I built pipelines to ingest and structure data from unstructured documents (invoices), integrating AI models to automate the extraction of key information. I also developed the backend of the web application used to visualize and manage the processed data.

OCR / NLPBACKENDDATA PIPELINES

Healthcare Data Architecture

At TDP Solutions, I designed the data architecture to centralize clinical records, integrating multiple hospital information sources for unified processing and visualization.

DATA ARCHITECTUREINTEGRATIONSQL

Tamborplast — Logistics Platform

At TDP Solutions, I built an internal full-stack web platform to coordinate shipments, logistics, and inventory management.

FULL STACKTYPESCRIPTNODE.JS

Datos al Ecosistema

National finalist (3rd place) for the Ministry of Mines and Energy: a data engineering solution with an automated ETL that consolidates heterogeneous natural-gas data from multiple entities (ANH, MME, UPME) for real-time analysis and decision making.

ETLPYTHONDATA ENGINEERING

Omegahack Hackathon 2025 — Winner

Winning project at the Grupo Nova / Universidad EAFIT hackathon: a real-time data analysis system that processes biometric and usage signals to detect behavioral patterns such as stress and demotivation.

REAL-TIMEPYTHONML

Legal AI Chatbot

At Universidad Nacional de Colombia, an AI chatbot built on local Hugging Face models that automatically extracts legal information from the university portal, letting users run complex queries over institutional regulations.

HUGGING FACENLPLOCAL LLM
Contact
© 2026 ALEJANDRO FERIA GONZALEZ
T R A C E T H E S C R E E N T O F L O W T H E I N K