Liresa Ferizaj, Senior Data Engineer and Team Lead
Liresa Ferizaj Senior Data Engineer | Team Lead

Databricks. Lakehouse. Data story.

Data engineered into decisions.

Liresa builds governed data platforms that turn scattered operational signals into trusted, decision-ready stories.

Follow the story CV available on request
0
faster onboarding
0
quality lift
0
daily pipeline success

The data story

From noise to narrative.

  1. 01 Source work starts at the contract

    Veeva, Oracle, SFTP, Fabric, survey, document, and media feeds get owners, cadence, schema expectations, and failure behavior before they land.

  2. 02 Governance is built into the lakehouse

    Delta Lake and Unity Catalog carry permissions, lineage, validation results, and audit context through Bronze, Silver, and Gold layers.

  3. 03 Pipelines make quality visible

    PySpark transformations, automated checks, profiling, and recovery paths turn raw movement into explainable data products teams can inspect.

  4. 04 Delivery stays close to users

    APIs, Databricks Apps, dashboards, and migration tooling package the platform into workflows people can actually use.

Data maturity line A minimal line chart rising from raw data to trusted decisions.

Working proof

Every source has an owner, cadence, schema expectation, and failure path. Intake discipline

Proof

The story holds because the systems hold.

0

to 60% faster source onboarding through configuration-driven ingestion.

0

better data quality through automated profiling, validation, and lineage.

0

less manual analysis through NLP and extraction workflows.

Selected work

Four chapters of range.

01

Healthcare lakehouse architecture

Databricks, Delta Lake, Unity Catalog, Workflows, Apps, ADF, and Snowflake.

02

AI-powered Alteryx to PySpark migration

Claude Sonnet, human approval, React, FastAPI, Databricks Apps, and Delta audit logs.

03

Humanitarian data products

ETL, PostgreSQL, MongoDB, FastAPI, survey data, documents, NLP, and extraction flows.

04

Computer vision for industrial quality

Python, OpenCV, NumPy, Seaborn, Pandas, and research-driven defect detection.

Expertise

A quiet stack for serious data work.

  • Databricks
  • Delta Lake
  • Unity Catalog
  • PySpark
  • Azure Data Factory
  • Snowflake
  • FastAPI
  • PostgreSQL
  • Data lineage
  • Data quality
  • AI migration
  • Team leadership

Field notes

Short essays on trust, migration, and meaning.

Read the blog

Contact

Build the next trusted data story.