Available for Q3/Q4 Consulting Engagements

High-Performance Cloud Data Platforms & AI-Ready Lakehouses

15+ years of enterprise data platform leadership. Architecting scalable AWS/GCP Medallion Lakehouses (Apache Iceberg, dbt), high-throughput PySpark pipelines, geospatial analytics, and automated data governance.

0 Years Data Engineering Leadership
0 Average Cloud Compute Cost Saved
0 Daily Events & Pipeline Records
Live Medallion Pipeline Monitor
AWS Iceberg + dbt
B
Bronze Layer (Raw Ingestion)
Streaming S3, EventBridge & Glue Crawlers
1.2 TB / day
S
Silver Layer (Cleaned & Normalized)
PySpark Star Schema & Schema Enforcement
-36.6% Storage
G
Gold Layer (AI-Ready Semantic)
dbt Iceberg Tables, Athena & Feature Store
< 120ms Query
Trusted by Enterprise Leaders & Innovation Teams
DB InfraGo
DB Fernverkehr AG
Telekom Deutschland
Zalando SE
Continental AG
Charly Education
Consulting Portfolio

End-to-End Data & Platform Engineering

From initial cloud data architecture blueprinting to hands-on execution and team mentoring.

AI-Ready Lakehouse & Modern Stack

Architecting robust, scalable Medallion data lakehouses (Bronze/Silver/Gold) supporting Analytics, Machine Learning, and GenAI feature store requirements with strict schema evolution.

Apache Iceberg AWS Glue dbt Terraform
☁️

Cloud Migration & Cost Optimization

Seamless migration from legacy Hadoop/Cloudera or Snowflake setups to native AWS/GCP cloud platforms. Refactoring ETL pipelines to cut infrastructure costs by up to 40%.

Cloudera to AWS Snowflake Refactoring PySpark EMR
🗺️

Geospatial Data Engineering

High-throughput processing of spatial datasets, vector geometries, and point clouds. Building distributed spatial indexing and sub-second GIS analytics queries.

PostGIS Trino GDAL/PDAL H3 Spatial Index
🛡️

Enterprise Data Governance & Security

Implementing AWS Lake Formation, Unity Catalog, and automated compliance enforcement agents (GDPR, Antitrust, PII masking, fine-grained access control).

AWS Lake Formation Unity Catalog GDPR Automation IAM / RBAC
🔄

High-Throughput ETL & Star Schema Optimization

Custom PySpark / Scala pipeline engineering, star-schema data volume reduction algorithms (extracting low-cardinality keys, surrogate integer mapping, Parquet compression).

Spark Tuning Star Schema Parquet Data Modeling
🎯

Fractional Lead Architect & Team Mentoring

Interim technical leadership, business information domain modeling, engineering best practices, CI/CD pipeline automation, and upskilling in-house data teams.

Domain Modeling Data Mesh CI/CD Technical Mentoring
Interactive Technical Demo

Data Architecture Playground

Inspect real implementation patterns, schema optimizations, and benchmarks used in enterprise consulting engagements.

Live Code & Metrics
AWS Apache Iceberg Medallion Lakehouse STATUS: OK
-- Code preview loading...
Proven Impact

Enterprise Case Studies

Click on any project to inspect the architecture challenge, implemented solution, and quantifiable results.

DB InfraGo / Deutsche Bahn

Enterprise Serverless Lakehouse & Governance

Designed and delivered an Apache Iceberg-based AWS lakehouse platform with granular RBAC governance and dbt semantic modeling.

35% Lower Cloud Costs & 99.9% Pipeline Reliability
Apache Iceberg AWS Glue dbt
DB Fernverkehr AG

Cloudera Hadoop to Native AWS Cloud Migration

Re-engineered legacy on-prem Hadoop workflows to native AWS Glue & EMR services with automated Infrastructure as Code.

🚀 3x Faster Batch Processing & Zero Hadoop Maintenance
PySpark AWS EMR Terraform
Telekom Deutschland GmbH

Snowflake to AWS Native Pipeline Optimization

Refactored heavy analytics queries and migrated workloads to native AWS Glue and Athena processing layers.

💰 40% Reduction in Monthly Analytics Expenditures
AWS Glue Athena PySpark
Zalando SE

GDPR & Data Governance Compliance Enforcement

Built automated compliance enforcement agent on Delta Lake to guarantee GDPR compliance and automated data masking across domains.

🔒 100% Automated Compliance Validation Across Lake Domains
Delta Lake Presto GDPR Agent
AI Data Platform Initiative

AI-Ready Medallion Blueprint Productization

Built reusable IaC product for analytics, ML, and GenAI feature engineering with embedded data quality frameworks.

🤖 Reduced GenAI Feature Prep from Weeks to Hours
Apache Iceberg dbt GenAI Ready
Sovereign Data Platform (OSS)

Self-Hosted Kubernetes Lakehouse & Open Catalog

Designed and built a fully open-source, self-hosted data platform running on Kubernetes (K3s) with Spark, Apache Polaris, Marimo, and Vault.

☸️ 100% Sovereign Data Ownership & Auto-scaling Compute
Kubernetes Apache Polaris Apache Spark Marimo
Ecosystem Expertise

Technology Stack & Tooling

Battle-tested enterprise technologies utilized across architecture engagements.

AWS
S3, Glue, Athena, EMR
Google Cloud
BigQuery, Dataproc
Apache Iceberg
ACID Table Format
Delta Lake
Lakehouse Storage
Apache Polaris
REST Metadata Catalog
Apache Spark
PySpark & Scala
dbt
Semantic Modeling
Marimo
Reactive Python Notebooks
Trino / Presto
Distributed SQL Engine
PostGIS
Geospatial Processing
Terraform
Infrastructure as Code
AWS CDK
Cloud Infrastructure
Kubernetes
Docker & ArgoCD
Dagster & Airflow
Orchestration & Lineage
Interactive Data Flow

Interactive Architecture Overview

Explore how enterprise data flows through the modern medallion architecture, semantic layers, processing compute engines, and downstream GenAI and BI applications. Hover over any node to trace lineage.

Architecture Pipeline Trace

Collaboration Framework

Consulting Engagement Models

Flexible consulting structures tailored to your data architecture maturity and team goals.

Architecture Audit & Strategy

1 - 2 Weeks Sprint

Rapid assessment of your existing cloud infrastructure, ETL bottlenecks, and cost drivers with actionable architectural recommendations.

  • Cloud Infrastructure & Cost Audit
  • ETL Pipeline Bottleneck Analysis
  • Security & Data Governance Check
  • Prioritized Architecture Roadmap
Request Audit

Fractional Lead Architect

Ongoing / Retainer Advisory

Embedded senior technical leadership guiding your data team through complex migrations, domain modeling, and high-scale scaling challenges.

  • Architectural Governance & Reviews
  • Hands-on PySpark/Iceberg Tuning
  • Data Mesh & Domain Modeling
  • Engineering Team Mentorship
Retain Lead Architect

Interactive Scope & Impact Estimator

Simulate the estimated timeline, expected ROI, and recommended engagement for your organization.

Estimated Project Duration
6 Weeks
Target Impact / ROI
4.0x Faster Queries
Recommended Service Package
Lakehouse Architecture & Implementation
Infrastructure Savings Projections
$100k $50k $0
$45,000
Current Cost
$29,000
Optimized
Background & Expertise

About Timor Ossevorth

Architecting Data Platforms That Scale with Precision

With over 15 years of hands-on data engineering experience, I specialize in transforming complex, siloed data environments into clean, high-throughput, AI-ready data platforms.

My career spans lead engineering roles at major enterprise organizations including Deutsche Bahn (DB InfraGo, DB Fernverkehr AG), Telekom Deutschland GmbH, and Zalando SE. I advocate for clean code, infrastructure as code, automated compliance, and open table formats like Apache Iceberg.

Education
B.Sc. Computer Engineering
Universität Potsdam (2008 - 2012)
Location
Berlin, Germany / Remote
Worldwide Enterprise Consulting
Languages
English (Fluent), German (Fluent)
Hebrew, Russian

Core Engineering Principles

  • 01 Domain Ownership & Data Mesh: Clear domain boundaries prevent monolithic bottlenecks.
  • 02 Open Table Formats: Avoid vendor lock-in by utilizing Apache Iceberg & Delta Lake.
  • 03 Infrastructure as Code: 100% reproducible environments via Terraform & AWS CDK.
  • 04 Cost Efficiency by Design: Optimize partition keys and column formats to minimize cloud bills.
Let's Connect

Ready to Upgrade Your Cloud Data Infrastructure?

Book an initial technical consultation to discuss your data architecture requirements, cloud migrations, or lakehouse initiatives.

✉️ Send Email to timor@timor-dataworks.com
📍 Berlin, Germany & Remote
📞 +49 163 230 1182
🌐 timor-dataworks.com
🔗 LinkedIn Profile