Skip to content
Available for new projects

Alonso Marcos Muñoz

Data Engineer · Pipelines, data modelling & platforms

I build data pipelines, models and platforms with Python and SQL. I bring professional public-metadata experience and applied projects with Databricks, Spark, Airflow, Kafka and AWS.

Python · SQL · PostgreSQL · Databricks · Spark · Airflow · Kafka · AWS

Albacete, Spain
Alonso Marcos Muñoz

About

Data Engineer focused on reliable pipelines, maintainable models and operable solutions. I work at Tragsatec within ImpulsaDATA for the Spanish Data Directorate, an initiative publicly reporting 5,771 published and federated datasets, 22 ministries and public bodies and 13 shared services.

My work combines Python, SQL and data modelling on PostgreSQL/Oracle with metadata ETL, CKAN, DCAT-AP-ES, RDF/JSON-LD and SHACL. It covers transformations, aggregations, views and optimisation alongside containerisation, security, test/dev/pre-production/production environments and operational documentation.

I complement this with a Master’s in Big Data & Cloud Computing, the CAPM certification and projects across OpenMetadata, Kubernetes, Airflow, Spark, Kafka, Databricks and AWS, plus applied generative AI and AI harness engineering to speed up analysis, development and documentation.

Professional experience

Tragsatec

TragsatecPresent

Oct 2025 – Present

Data Engineer · ImpulsaDATA (Spanish Data Directorate)

Public sector · Spain

datos.gob.es

Data engineering and data governance in Spain’s central government. ImpulsaDATA publicly reports 5,771 published and federated datasets, 22 ministries and public bodies and 13 shared services.

  • Evolved and maintained metadata ETL pipelines: normalisation, mapping to common models and publishing into CKAN.
  • Aligned metadata with DCAT-AP-ES and produced RDF and JSON-LD outputs with SHACL semantic validation before and after catalogue ingestion.
  • Built and evolved CKAN extensions and plugins (DCAT, scheming, harvest, spatial) plus a custom plugin shared across all public bodies.
  • Configured containerised stacks (Docker/Podman) and administered the team’s test, development, pre-production and production environments.
  • Modelled data in SQL over the metadata model (transformations, aggregations, views and optimisation) on PostgreSQL.
  • Wrote operational documentation and mentored junior profiles, with a focus on quality, maintainability and low technical debt.
Tragsatec

Tragsatec

Jun 2025 – Sep 2025

Data Engineer Intern · ImpulsaDATA

Province of Albacete · Spain

datos.gob.es

First stage in the ImpulsaDATA project, focused on the initial design and development of a metadata ETL converter for public-sector data-governance workflows.

  • Developed metadata ETL flows in Python, from extraction and normalisation to standardised JSON/JSON-LD structures.
  • Mapped metadata to a common DCAT-AP-ES-aligned model and generated RDF for interoperable publishing.
  • Performed SHACL semantic validation and integrated with CKAN catalogues through the CKAN Action API.
  • Implemented external configuration, structured logs, technical documentation, diagrams, data dictionaries and operating guides.
La Fábrica del Tiempo

La Fábrica del Tiempo

Feb 2024 – Aug 2024

Process Automation & Productivity Consultant

Almansa, Spain

Process automation and productivity with Power Platform on Microsoft 365, including client training and technical content creation.

  • Internal app and no-code flows with Power Apps + Power Automate for CRM lead management.
  • SharePoint–CRM synchronisation and automated notifications/approvals.
  • Over 25% improvement in operational efficiency on top of the ADOCC methodology.
Ayuntamiento de Alcázar de San Juan

Ayuntamiento de Alcázar de San Juan

Mar 2019 – Jun 2019

IT Technician

Local government · Spain

IT support and basic systems & network administration in a municipal setting.

  • Incident resolution and end-user support.
  • Systems maintenance and basic infrastructure tasks.
  • Technical documentation within a public administration context.
Soporte ITSistemasRedes

Applied Big Data and cloud

Applied academic projects with architecture, execution evidence and limitations

Visual evidence for Telco churn Lakehouse and MLOps
Big Data · MLOps2026
Applied academic project

Telco churn Lakehouse and MLOps

Reproducible Databricks Medallion Lakehouse and MLOps lifecycle with execution evidence.

DatabricksDelta LakeMLflowUnity Catalog
View case study
Visual evidence for Smart Parking Albacete
Cloud · IoT2026
Applied academic project

Smart Parking Albacete

Serverless IoT platform tested in AWS Academy, from MQTT telemetry to API and dashboard.

AWS IoTLambdaDynamoDBStreamlit
View case study
Visual evidence for Spark, Kafka and Airflow data platform
Data Engineering2026
Applied academic project

Spark, Kafka and Airflow data platform

Reproducible local platform with batch, streaming, Medallion architecture and orchestration.

SparkAirflowKafkaDelta Lake
View case study

Tech stack

Data governance & quality

Applied AI & quality

AI harness engineeringContext engineeringAutomatización de flujos técnicosPruebas end-to-end asistidasRevisión técnica de códigoDocumentación reproducible

Management & languages

PMI / CAPMKanbanADOCCEspañol (nativo)Inglés B2

Education & certifications

UCLM
Education

Master's Degree in Big Data & Cloud Computing

UCLM · Sep 2025 – Jun 2026

Completed · Master's thesis 9.2/10
UCLM
Education

BSc in Computer Engineering · Information Technologies specialisation

UCLM · Sep 2019 – Jun 2025

Education

Higher Vocational Diploma in Network Systems Administration

IES Juan Bosco · Sep 2017 – Jun 2019

Project Management Institute
Certification

CAPM

Project Management Institute

Certified foundations of PMI project management.
Cambridge English
Certification

First Certificate in English (B2)

Cambridge English

Accredited professional English proficiency.

Let's talk

Open to opportunities in data engineering, data governance and data platforms.

Live GitHub signal

Coding stats

Public activity refreshed from GitHub on every deployment.

Open GitHub profile

Alonso Marcos Muñoz's GitHub

Total commits (2026)
179
Total PRs
20
Merged PRs
100.0%
Total issues
56
Contributions (last 12 months)
258
100%merged

Most used languages

Share of code bytes across owned public repositories.

  • Jupyter Notebook26.7%
  • Python25.0%
  • TeX20.0%
  • TypeScript10.5%
  • HTML7.9%
  • Astro3.8%

Public activity onlyUpdated 14 Sept 2026GitHub API