Kiz8
All open roles
/ CAREERS

Middle / Senior Data Engineer

Build and operate the lakehouse, pipelines, and data marts that power analytics and AI agent workflows.

About Kiz8

Kiz8 helps enterprises and high-growth companies deploy AI agent ecosystems and modern data analytics solutions. We turn complex, fragmented operations into streamlined, data-driven workflows. Our data infrastructure is built around a modern lakehouse architecture using open table formats, fast analytical engines, and scalable serverless and edge components.

The role

We are looking for a Middle or Senior Data Engineer who understands modern data architectures and has hands-on experience making them reliable in production. You will own the design, scaling, and monitoring of end-to-end data pipelines, analytical data marts, and lakehouse layers used by internal BI analytics and autonomous AI agent workflows.

What you will own

Lakehouse architecture

Design, maintain, and optimize storage layers with Apache Iceberg and Parquet, including partitioning, compaction, and schema evolution.

Pipeline development and monitoring

Build and monitor resilient batch and real-time ETL and ELT pipelines across a range of data sources.

Data marts and analytical engines

Model and build high-performance data marts and staging layers in ClickHouse and DuckDB.

Operational database integration

Manage ingestion, synchronization, and CDC or replication workflows from PostgreSQL.

Edge and serverless infrastructure

Integrate data processes with Cloudflare Workers, R2, Queues, and related services for distributed, efficient pipelines.

Data quality and reliability

Set up automated tests, lineage tracking, monitoring, and alerts to maintain accuracy and availability.

Collaboration with AI teams

Partner with AI engineers and analysts to deliver structured data layers, features, and context datasets for AI agents.

What you bring

Production experience

3+ years for Middle or 5+ years for Senior in data engineering, analytics engineering, or backend data infrastructure.

Lakehouse and formats

Practical knowledge of Apache Iceberg, Parquet, and object storage patterns using Cloudflare R2 or S3.

Analytical engines and databases

Strong experience with ClickHouse and DuckDB, plus deep operational proficiency with PostgreSQL and advanced SQL.

Programming

Solid software engineering skills in Python and/or TypeScript with Node.js.

Data modeling

A clear understanding of dimensional modeling, data mart design, and data quality frameworks.

Infrastructure mindset

Comfort working with Docker and modern serverless or edge environments.

Nice to have

  • Experience with distributed SQL or streaming query engines such as Trino or Apache Flink.
  • Hands-on experience with CDC tools such as Debezium or pglogical.
  • A background in preparing data for LLMs, vector search, or AI agent architectures.
  • Experience with Dagster, Airflow, Temporal, or another workflow orchestration engine.

Location and work setup

Fully remote

We are a remote-first team and welcome applications from candidates regardless of location.

Lisbon advantage

Being based in or near Lisbon is a strong plus and creates opportunities for in-person collaboration, meetups, and strategy sessions.

What we offer

Direct ownership

High autonomy to build core architecture without legacy enterprise red tape.

Modern stack

Work with Iceberg, DuckDB, ClickHouse, Cloudflare, and AI agents instead of outdated data platforms.

Flexible environment

A remote-friendly culture focused on results, autonomy, and sustainable work.

Competitive compensation

A package aligned with your experience level and the impact you can make.

Apply for this role

To apply, email your CV to sean@kiz8.team. We review every application and will contact you if your experience matches the role.

Email your CV