# Senior Data Engineer

[Huskeys](https://gurify.com/jobs?q=Huskeys) · Tel Aviv · Posted last week

Senior

[Data](https://gurify.com/jobs/data)

[Apply on the original posting → (opens in a new tab)](https://jobs.ashbyhq.com/huskeys/c1a573c1-7ea2-43b1-a94b-12991562d2be)

## Job description

Huskeys is reimagining web application protection for the modern era. As applications become increasingly dynamic, multi-cloud, and AI-driven, traditional WAFs struggle to keep up.

We built the first Agentic Network Security Control Plane - an AI-powered mitigation layer that works with any existing WAF. Our platform helps security teams unify protection, measure effectiveness, reduce manual overhead, and create context-aware policies that block threats without disrupting business flows.

At Huskeys, we believe context is the new attack surface, and that security should empower the business - not slow it down.

Security runs on data. Every decision our platform makes - every threat we catch and every business flow we keep open - depends on our ability to ingest, partition, aggregate, and trust massive volumes of log and scan data. This role owns that foundation.

Why we need you?

Our platform is fundamentally data-driven: we ingest huge volumes of logs, scans, and telemetry, and the value we deliver depends on how well we partition, aggregate, and cross-reference all of it. Today this work is spread thin, and we need a dedicated owner for it.

You'll own the scale and performance of our data layer - designing partitioning and aggregation strategies that hold up as log volume and complexity grow, rather than ad-hoc queries that fall over. You'll be the guardian of data integrity and trust, building reliable ETLs from many different sources (cloud providers, outside scans) and the consistency guarantees that let the whole team build with confidence. And you'll set the stones for ML: we're not building models yet, but we want our data structured, clean, and accessible so we can add ML without re-plumbing everything.

Without this role, our data work stays fragmented, aggregations get slower and less trustworthy as we scale, and we stay blocked from making the most of the data we already have.

### About the position

At Huskeys, our platform ingests and cross-references massive volumes of logs, scans, and telemetry from cloud providers, CDNs/WAFs, and external scanning tools. The depth and reliability of our data layer is what makes our context-aware security decisions possible.

We are looking for a Data Engineer who can independently own that data layer end-to-end - from table and schema design through partitioning, complex aggregation, and multi-source ETL - and who thinks beyond "can I write this query?" to "will this scale, is the data correct, and can the team build on it?"

### What Success Looks Like After 6 Months

- Independent Owner: You independently own log analysis and partitioning, our complex aggregations and cross-referencing, and the strategy for scaling our data layer - making architectural calls without constant oversight.

- Source of Truth: Our internal data is the best-in-class source of truth across every source we touch - cloud provider and CDN/WAF scans, log analysis, and external scanning - with strong, shared definitions for how we model, unify, and aggregate data.

- Scalable Foundation: Adding new complex aggregations is easy and safe at our scale, with best practices built in, and the team has a data strategy everyone can build on.

- Business Enabler: Our data layer runs with high integrity and strong SLAs, scales to onboard more clients and new use cases, and deepens our understanding of customers' environments - keeping the product growing in a scalable direction.

### Key Responsibilities

- Table & Schema Structure - Design and own our table structures - data models, keys, types, partitioning, and relationships - and the unified models that turn many sources into one coherent, queryable shape.

- Log Analysis & Partitioning - Design and own how we store and partition large-scale log data so it stays fast and queryable as volume grows.

- Complex Aggregation & Cross-Referencing - Build and optimize the complex aggregations and cross-source joins that power our platform's insights and decisions.

- ETLs from Multiple Sources - Build and maintain reliable ETL/ELT pipelines from many different sources - cloud providers and outside scans - landing diverse inputs cleanly into our unified tables.

- Data Integrity & Observability - Ensure data consistency and correctness, and monitor pipeline health, freshness, and quality with alerting, so we meet our SLAs and catch bad data before it reaches the product.

- Data Governance - Handle sensitive customer data responsibly - access control, retention, and privacy - in line with our security and compliance commitments.

- Scaling Strategy - Own the strategy for scaling our data and aggregations - table design, partitioning, query performance, and cost - ahead of where we are today.

- Setting the Stones for ML - Structure and prepare data (clean, feature-ready tables and pipelines) so we're ready to add machine learning when the business is.

### Day-to-day Reality

- Types of Tasks: Primarily hands-on data engineering and system design - schema and partitioning design, building pipelines, writing and optimizing complex queries. Some meetings for planning and collaboration, but at least 70-80% is technical work.

- Level of Independence: High autonomy with end-to-end ownership. When you take on the data layer, it's yours A to Z - design, implementation, monitoring, and long-term maintenance. You make architectural decisions independently with team input.

- Pace: Fast-paced and dynamic. Seed-stage startup where priorities can shift; you balance building new data capabilities with keeping existing pipelines reliable.

- Collaboration: A mix of independent work and cross-team collaboration. You'll work closely with backend engineers, product, and, over time, anyone building on top of our data.

- Tools/Environment: Python, SQL, ClickHouse (columnar/log analysis), PostgreSQL (NeonDB), Temporal for workflow orchestration, AWS Athena, and modern observability platforms (Datadog/Grafana/ELK). We're flexible on exact tools - equivalent experience counts.

### The Ideal candidate

- You are a pragmatic builder who balances quality and speed - you know when to ship a workable pipeline and when to invest in robust, scalable data architecture.

- You are a systems thinker who designs data models, partitioning, and aggregation strategies that hold up as volume grows.

- You have an ownership mindset - you take the data layer A to Z without hand-holding, thinking beyond "can I write this query?" to "will this scale, is the data correct, and can the team build on it?"

- You are a collaborative engineer who enables teammates through clean, well-modeled, well-documented datasets and knowledge sharing.

- You are a strong communicator who can explain data trade-offs to both engineers and non-technical stakeholders.

- You thrive in startup environments with ambiguity and rapid change, and you care deeply about data integrity and correctness - not just getting a number out the door.

- You're comfortable making decisions independently while seeking input when needed, and you fit the HUSKEY values: accountable, driven, adaptable, sharp, and positive.

### Must-Have Requirements

- 5+ years of data engineering (or strongly data-focused backend) experience.

- Expert-level SQL - designing schemas, modeling data, and writing and optimizing complex aggregation and cross-referencing queries.

- Strong Python for building ETL/ELT pipelines and data tooling.

- Large-scale data experience - hands-on with high-volume log/event data, including partitioning and performance optimization (ClickHouse, Athena, or similar columnar/analytical stores).

- ETL/pipeline experience - has built and maintained reliable pipelines ingesting from multiple, diverse sources.

- Data integrity & consistency - solid understanding of data quality, consistency, and reliability across distributed environments.

- Scaling mindset - can design data and aggregation strategies that scale, not just work today.

Note: We're flexible on specific tools. If a candidate has done equivalent work with comparable technologies, that counts - we don't require our exact stack.

**Live in Huskeys’s hiring system.** Read from the company's own applicant tracking system, not reposted from a job board — so it's a real, open requisition rather than an ad that outlived the role.

We remove it as soon as it disappears at source.

## More jobs like this

- QU [Data Engineer](https://gurify.com/job/data-engineer-at-quicklizard-1abe5e1cf1a6) Quicklizard · Petah Tikva, Israel · 4 days ago
- SI [Data Engineer – Applied ML](https://gurify.com/job/data-engineer-applied-ml-at-similarweb-bc9ef18b9695) Similarweb · Tel Aviv, Israel · 5 days ago
- TR [Data Engineer](https://gurify.com/job/data-engineer-at-trigo-ca95ed09ab0a) Trigo · Ramat Gan, Israel · 3 weeks ago
- VE [Data Lead- Full-Stack Data Engineer](https://gurify.com/job/data-lead-full-stack-data-engineer-at-venncity-02ec7b8acca3) Venncity · Tel Aviv District, Israel · 3 weeks ago
- CI [Analytics Data Engineer](https://gurify.com/job/analytics-data-engineer-at-comm-it-9f06680c7994) Comm IT · Israel · 3 weeks ago
- AT [Senior BI Engineer](https://gurify.com/job/senior-bi-engineer-at-atbay-d2ea8520e9e3) Atbay · Tel Aviv-Yafo, Israel · 5 days ago

```json
{"@context":"https://schema.org/","@type":"JobPosting","title":"Senior Data Engineer","description":"\u003Cp\u003EHuskeys is reimagining web application protection for the modern era. As applications become increasingly dynamic, multi-cloud, and AI-driven, traditional WAFs struggle to keep up.\u003C/p\u003E\u003Cp\u003EWe built the first Agentic Network Security Control Plane - an AI-powered mitigation layer that works with any existing WAF. Our platform helps security teams unify protection, measure effectiveness, reduce manual overhead, and create context-aware policies that block threats without disrupting business flows.\u003C/p\u003E\u003Cp\u003EAt Huskeys, we believe context is the new attack surface, and that security should empower the business - not slow it down.\u003C/p\u003E\u003Cp\u003ESecurity runs on data. Every decision our platform makes - every threat we catch and every business flow we keep open - depends on our ability to ingest, partition, aggregate, and trust massive volumes of log and scan data. This role owns that foundation.\u003C/p\u003E\u003Cp\u003EWhy we need you?\u003C/p\u003E\u003Cp\u003EOur platform is fundamentally data-driven: we ingest huge volumes of logs, scans, and telemetry, and the value we deliver depends on how well we partition, aggregate, and cross-reference all of it. Today this work is spread thin, and we need a dedicated owner for it.\u003C/p\u003E\u003Cp\u003EYou\u0026#39;ll own the scale and performance of our data layer - designing partitioning and aggregation strategies that hold up as log volume and complexity grow, rather than ad-hoc queries that fall over. You\u0026#39;ll be the guardian of data integrity and trust, building reliable ETLs from many different sources (cloud providers, outside scans) and the consistency guarantees that let the whole team build with confidence. And you\u0026#39;ll set the stones for ML: we\u0026#39;re not building models yet, but we want our data structured, clean, and accessible so we can add ML without re-plumbing everything.\u003C/p\u003E\u003Cp\u003EWithout this role, our data work stays fragmented, aggregations get slower and less trustworthy as we scale, and we stay blocked from making the most of the data we already have.\u003C/p\u003E\u003Ch3\u003EAbout the position\u003C/h3\u003E\u003Cp\u003EAt Huskeys, our platform ingests and cross-references massive volumes of logs, scans, and telemetry from cloud providers, CDNs/WAFs, and external scanning tools. The depth and reliability of our data layer is what makes our context-aware security decisions possible.\u003C/p\u003E\u003Cp\u003EWe are looking for a Data Engineer who can independently own that data layer end-to-end - from table and schema design through partitioning, complex aggregation, and multi-source ETL - and who thinks beyond \u0026quot;can I write this query?\u0026quot; to \u0026quot;will this scale, is the data correct, and can the team build on it?\u0026quot;\u003C/p\u003E\u003Ch3\u003EWhat Success Looks Like After 6 Months\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EIndependent Owner: You independently own log analysis and partitioning, our complex aggregations and cross-referencing, and the strategy for scaling our data layer - making architectural calls without constant oversight.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ESource of Truth: Our internal data is the best-in-class source of truth across every source we touch - cloud provider and CDN/WAF scans, log analysis, and external scanning - with strong, shared definitions for how we model, unify, and aggregate data.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EScalable Foundation: Adding new complex aggregations is easy and safe at our scale, with best practices built in, and the team has a data strategy everyone can build on.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EBusiness Enabler: Our data layer runs with high integrity and strong SLAs, scales to onboard more clients and new use cases, and deepens our understanding of customers\u0026#39; environments - keeping the product growing in a scalable direction.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EKey Responsibilities\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003ETable \u0026amp; Schema Structure - Design and own our table structures - data models, keys, types, partitioning, and relationships - and the unified models that turn many sources into one coherent, queryable shape.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ELog Analysis \u0026amp; Partitioning - Design and own how we store and partition large-scale log data so it stays fast and queryable as volume grows.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EComplex Aggregation \u0026amp; Cross-Referencing - Build and optimize the complex aggregations and cross-source joins that power our platform\u0026#39;s insights and decisions.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EETLs from Multiple Sources - Build and maintain reliable ETL/ELT pipelines from many different sources - cloud providers and outside scans - landing diverse inputs cleanly into our unified tables.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EData Integrity \u0026amp; Observability - Ensure data consistency and correctness, and monitor pipeline health, freshness, and quality with alerting, so we meet our SLAs and catch bad data before it reaches the product.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EData Governance - Handle sensitive customer data responsibly - access control, retention, and privacy - in line with our security and compliance commitments.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EScaling Strategy - Own the strategy for scaling our data and aggregations - table design, partitioning, query performance, and cost - ahead of where we are today.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ESetting the Stones for ML - Structure and prepare data (clean, feature-ready tables and pipelines) so we\u0026#39;re ready to add machine learning when the business is.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EDay-to-day Reality\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003ETypes of Tasks: Primarily hands-on data engineering and system design - schema and partitioning design, building pipelines, writing and optimizing complex queries. Some meetings for planning and collaboration, but at least 70-80% is technical work.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ELevel of Independence: High autonomy with end-to-end ownership. When you take on the data layer, it\u0026#39;s yours A to Z - design, implementation, monitoring, and long-term maintenance. You make architectural decisions independently with team input.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EPace: Fast-paced and dynamic. Seed-stage startup where priorities can shift; you balance building new data capabilities with keeping existing pipelines reliable.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ECollaboration: A mix of independent work and cross-team collaboration. You\u0026#39;ll work closely with backend engineers, product, and, over time, anyone building on top of our data.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ETools/Environment: Python, SQL, ClickHouse (columnar/log analysis), PostgreSQL (NeonDB), Temporal for workflow orchestration, AWS Athena, and modern observability platforms (Datadog/Grafana/ELK). We\u0026#39;re flexible on exact tools - equivalent experience counts.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EThe Ideal candidate\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EYou are a pragmatic builder who balances quality and speed - you know when to ship a workable pipeline and when to invest in robust, scalable data architecture.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EYou are a systems thinker who designs data models, partitioning, and aggregation strategies that hold up as volume grows.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EYou have an ownership mindset - you take the data layer A to Z without hand-holding, thinking beyond \u0026quot;can I write this query?\u0026quot; to \u0026quot;will this scale, is the data correct, and can the team build on it?\u0026quot;\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EYou are a collaborative engineer who enables teammates through clean, well-modeled, well-documented datasets and knowledge sharing.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EYou are a strong communicator who can explain data trade-offs to both engineers and non-technical stakeholders.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EYou thrive in startup environments with ambiguity and rapid change, and you care deeply about data integrity and correctness - not just getting a number out the door.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EYou\u0026#39;re comfortable making decisions independently while seeking input when needed, and you fit the HUSKEY values: accountable, driven, adaptable, sharp, and positive.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EMust-Have Requirements\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003E5\u002B years of data engineering (or strongly data-focused backend) experience.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EExpert-level SQL - designing schemas, modeling data, and writing and optimizing complex aggregation and cross-referencing queries.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EStrong Python for building ETL/ELT pipelines and data tooling.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ELarge-scale data experience - hands-on with high-volume log/event data, including partitioning and performance optimization (ClickHouse, Athena, or similar columnar/analytical stores).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EETL/pipeline experience - has built and maintained reliable pipelines ingesting from multiple, diverse sources.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EData integrity \u0026amp; consistency - solid understanding of data quality, consistency, and reliability across distributed environments.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EScaling mindset - can design data and aggregation strategies that scale, not just work today.\u003C/li\u003E\u003C/ul\u003E\u003Cp\u003ENote: We\u0026#39;re flexible on specific tools. If a candidate has done equivalent work with comparable technologies, that counts - we don\u0026#39;t require our exact stack.\u003C/p\u003E","identifier":{"@type":"PropertyValue","name":"Gurify","value":"senior-data-engineer-at-huskeys-75f4560f78ed"},"url":"https://gurify.com/job/senior-data-engineer-at-huskeys-75f4560f78ed","datePosted":"2026-09-22","validThrough":"2026-11-19T23:59:59Z","hiringOrganization":{"@type":"Organization","name":"Huskeys","sameAs":"https://jobs.ashbyhq.com/huskeys"},"directApply":false,"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressCountry":"IL","addressLocality":"Tel Aviv"}}}
```

```json
{"@context":"https://schema.org/","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Jobs","item":"https://gurify.com/jobs"},{"@type":"ListItem","position":2,"name":"Israel","item":"https://gurify.com/jobs/israel"},{"@type":"ListItem","position":3,"name":"Senior Data Engineer","item":"https://gurify.com/job/senior-data-engineer-at-huskeys-75f4560f78ed"}]}
```
