TUM M.Sc. Informatics • Cloud, Data & AI Engineering

Carlos Mejia

Cloud & Data Engineer & Applied AI Researcher

Bridging rigorous academic AI research at Technical University of Munich with 9+ years of enterprise software and data platform architecture across AWS and GCP. Focused on production infrastructure-as-code, multi-environment data pipelines, MLOps, and agentic workflows.

🏛️ Academic TUM Munich M.Sc. Informatics & AI Systems
⚡ Experience 9+ Years Software & Data Platforms
☁️ Multi-Cloud AWS & GCP Terraform, Docker & CI/CD
📍 Location Munich, Germany Hybrid & Global Mobility

About

Carlos has more than nine years of experience in software and data development, combining consultancy, startups, and large companies across mobility, logistics, media, and healthcare. He has led teams of 2 to 3 people and is currently expanding his technical experience and knowledge with a master's degree in Informatics at TUM Munich, focused on cloud computing, machine learning, including deep learning and natural language processing, and software engineering.

He is equally comfortable in the fast-paced environment of a startup, where he has designed data architectures from scratch, and in the more structured environment of a large company, where he has built production infrastructure at scale. That combination allows him to bring strong engineering practices to small teams without slowing down the speed they need.

He is excited by complex technical and business challenges related to collecting, interpreting, analyzing, and modeling data, and increasingly with orchestrating AI agents in real engineering workflows. He also enjoys sharing his knowledge, as he has taught classes at both personal and university levels on topics such as machine learning, MLOps, data engineering, and data analysis.

Interactive Skill Matrix

Click any skill to highlight related career milestones, architecture projects, and research artifacts across the site.

Showing entries for selected skill

☁️ Cloud & DevOps

📊 Data Engineering

🧠 AI & Software

Career & Research Timeline

BMW Group • Cloud DevOps Engineer Oct 2024 – Present (Munich)
Enterprise

Maintain AWS infrastructure-as-code with Terraform and AWS CDK across production mobility services; orchestrate batch data pipelines with EMR/Spark, AWS Glue, and Python; configure IAM and KMS security hardening.

AWSTerraformSparkGluePythonIAM
Salsuki • Solo Full-Stack Engineer 2024 – Present (Munich)
Startup / Community

Built and maintain a full-stack platform for a non-profit salsa community (React 19, TypeScript, Supabase, Postgres RLS, Deno Edge Functions). Orchestrated using an AI-agent-driven solo development workflow.

ReactTypeScriptSupabasePostgreSQLGitHub Actions
Ocumeda GmbH • Data Engineer May 2024 – Sep 2024 (Munich)
HealthTech Startup

Designed scalable AWS serverless data architecture (S3, Lambda, Redshift, Step Functions); developed ~10 ETL pipelines with Python connecting clinical APIs with up to 400% execution speedups.

AWSPythonPostgreSQLETL
TUM Robotics & AI Chair • M.Sc. Thesis Researcher 2024 – Present (Munich)
Academic Research

Knowledge Graph Construction & Retrieval for LLM-Based Test Scenario Generation from UN Regulation 152. Benchmarking RegulaRAG, LightRAG, and HippoRAG 2.

Knowledge GraphsPythonPyTorchLangChainWeaviate
Delta Smith • Data Engineer & Architect (Freelance) Sep 2021 – Sep 2023 (Mexico City)
Freelance Consulting

Architected batch data pipelines and observability infrastructure with Google Cloud Composer (Airflow), BigQuery, Docker, and Terraform for commercial clients.

GCPTerraformAirflowDockerPostgreSQL
ITESM • University Lecturer (MLOps & Data Engineering) Feb 2022 – Jun 2023 (Mexico City)
University Academia

Taught courses in the Master's in Applied AI and MLOps certification: Python OOP, Pytest, FastAPI, Docker, Terraform, GitHub Actions, and Observability.

PythonFastAPIDockerTerraformGitHub Actions
Wizeline • Data Engineer Jan 2022 – Apr 2023 (Mexico City)
Enterprise Clients

Developed streaming and batch ETL pipelines across GCP and AWS for enterprise clients (NBCUniversal, Dow Jones); built FastAPI ML model microservices on Docker and AWS EC2; taught Terraform and Kafka.

GCPAWSKafkaSparkFastAPITerraform
iVoy • Data Engineer Feb 2021 – Jan 2022 (Mexico City)
Logistics Startup

Consolidated MySQL, Postgres, and MongoDB streams into a Kafka cluster (500M+ records); built streaming pipelines with Apache Beam on Dataflow and an AWS S3 data lake with Glue and Athena.

Apache BeamKafkaSparkAirflowPostgreSQLAWS Glue
Uber Technologies Inc • Operations & Automation Intern Oct 2019 – Jan 2021 (Mexico City)
Enterprise

Productized a Python GUI automation app saving 40+ hours/month; built document classification prototypes using Deep Learning (ResNet-50) and OCR.

PythonPyTorchSQLOCR

Featured Projects

Salsuki: Community Management Platform

2024 – Current

Full-stack platform built on Supabase (Postgres with RLS, Edge Functions, Auth, Storage) and React 19/TypeScript on Vercel with an append-only ledger economy and automated CI/CD.

Read the full architecture write-up →

LECture-bot: GenAI Course Assistant

Apr 2025 – Aug 2025

Generative AI and RAG subsystem built with LangChain, Weaviate vector storage, and local/cloud LLMs for course material intelligence.

Read the full project write-up →

Education

Technical University of Munich

Oct 2023 – Expected Apr 2027

M.Sc. in Informatics focused on distributed cloud computing, machine learning, and reliable software systems.

Instituto Politécnico Nacional (UPIITA)

Jan 2015 – Jul 2020

B.Sc. in Telematics Engineering.

Master's Thesis

KG Construction and Retrieval for LLM-Based Test Scenario Generation from Automotive Regulations

Implementation underway (Munich)

TUM Chair of Robotics, AI and Real-Time Systems, advised by Prof. Alois Knoll and André Schamschurko.

Designing and evaluating a structured Knowledge Graph retrieval layer to improve LLM-based generation of test scenarios (speed / load condition / post-condition tuples) from automotive regulations (UN Regulation 152). Benchmarked against three systems: RegulaRAG, LightRAG, and HippoRAG 2, using precision, recall, F1, and success or failure rate metrics.

Read the research overview →