# Senior Solutions Engineer – Big Data & Data Infrastructure

[Vastdata](https://gurify.com/jobs?q=Vastdata) · Tel Aviv-Yafo, Israel · Posted 2 weeks ago

Senior

[Apply on the original posting → (opens in a new tab)](https://www.comeet.com/jobs/vastdata/43.001/senior-solutions-engineer--big-data--data-infrastructure/BF.546)

## Job description

This is a great opportunity to be part of one of the fastest-growing infrastructure companies in history, an organization that is in the center of the hurricane being created by the revolution in artificial intelligence.

"VAST's data management vision is the future of the market."- Forbes

VAST Data is the data platform company for the AI era. We are building the enterprise software infrastructure to capture, catalog, refine, enrich, and protect massive datasets and make them available for real-time data analysis and AI training and inference. Designed from the ground up to make AI simple to deploy and manage, VAST takes the cost and complexity out of deploying enterprise and AI infrastructure across data center, edge, and cloud.

Our success has been built through intense innovation, a customer-first mentality and a team of fearless VASTronauts who leverage their skills & experiences to make real market impact. This is an opportunity to be a key contributor at a pivotal time in our company’s growth and at a pivotal point in computing history.

We are seeking an experienced Solutions Data Engineer who possess both technical depth and strong interpersonal skills to partner with internal and external teams to develop scalable, flexible, and cutting-edge solutions. Solutions Engineers collaborate with operations and business development to help craft solutions to meet customer business problems.

A Solutions Engineer works to balance various aspects of the project, from safety to design. Additionally, a Solutions Engineer researches advanced technology regarding best practices in the field and seek to find cost-effective solutions.

### Job Description:

We’re looking for a Solutions Engineer with deep experience in Big Data technologies, real-time data pipelines, and scalable infrastructure—someone who’s been delivering critical systems under pressure, and knows what it takes to bring complex data architectures to life. This isn’t just about checking boxes on tech stacks—it’s about solving real-world data problems, collaborating with smart people, and building robust, future-proof solutions.

In this role, you’ll partner closely with engineering, product, and customers to design and deliver high-impact systems that move, transform, and serve data at scale. You’ll help customers architect pipelines that are not only performant and cost-efficient but also easy to operate and evolve.

We want someone who’s comfortable switching hats between low-level debugging, high-level architecture, and communicating clearly with stakeholders of all technical levels.

### Key Responsibilities:

- Build distributed data pipelines using technologies like Kafka, Spark (batch & streaming), Python, Trino, Airflow, and S3-compatible data lakes—designed for scale, modularity, and seamless integration across real-time and batch workloads.

- Design, deploy, and troubleshoot hybrid cloud/on-prem environments using Terraform, Docker, Kubernetes, and CI/CD automation tools.

- Implement event-driven and serverless workflows with precise control over latency, throughput, and fault tolerance trade-offs.

- Create technical guides, architecture docs, and demo pipelines to support onboarding, evangelize best practices, and accelerate adoption across engineering, product, and customer-facing teams.

- Integrate data validation, observability tools, and governance directly into the pipeline lifecycle.

- Own end-to-end platform lifecycle: ingestion → transformation → storage (Parquet/ORC on S3) → compute layer (Trino/Spark).

- Benchmark and tune storage backends (S3/NFS/SMB) and compute layers for throughput, latency, and scalability using production datasets.

- Work cross-functionally with R&D to push performance limits across interactive, streaming, and ML-ready analytics workloads.

- Operate and debug object store–backed data lake infrastructure, enabling schema-on-read access, high-throughput ingestion, advanced searching strategies, and performance tuning for large-scale workloads.

### Required Skills & Experience:

- 2–4 years in software / solution or infrastructure engineering, with 2–4 years focused on building / maintaining large-scale data pipelines / storage & database solutions.

- Proficiency in Trino, Spark (Structured Streaming & batch) and solid working knowledge of Apache Kafka.

- Coding background in Python (must-have); familiarity with Bash and scripting tools is a plus.

- Deep understanding of data storage architectures including SQL, NoSQL, and HDFS.

- Solid grasp of DevOps practices, including containerization (Docker), orchestration (Kubernetes), and infrastructure provisioning (Terraform).

- Experience with distributed systems, stream processing, and event-driven architecture.

- Hands-on familiarity with benchmarking and performance profiling for storage systems, databases, and analytics engines.

- Excellent communication skills—you’ll be expected to explain your thinking clearly, guide customer conversations, and collaborate across engineering and product teams.

**Live in Vastdata’s hiring system.** Read from the company's own applicant tracking system, not reposted from a job board — so it's a real, open requisition rather than an ad that outlived the role.

We remove it as soon as it disappears at source.

## More jobs like this

- VA [Customer Success Engineer at VAST Data](https://gurify.com/job/customer-success-engineer-at-vast-data-at-vastdata-ed2c8915ed0d) Vastdata · Tel Aviv-Yafo, Israel · 3 weeks ago
- SE [Senior Data Engineer](https://gurify.com/job/senior-data-engineer-at-seekingalpha-34ab8c585894) Seekingalpha · Ra'anana, Israel · 3 days ago
- SO [Data Engineer](https://gurify.com/job/data-engineer-at-somekhchaikin-0165753a5520) Somekhchaikin · Tel Aviv-Yafo, Israel · 5 days ago
- GU [Data Engineer](https://gurify.com/job/data-engineer-at-guardio-c214b8fd240a) Guardio · Tel Aviv-Yafo, Israel · last week
- TE [Reindeer- Software Engineer (Infrastructure)](https://gurify.com/job/reindeer-software-engineer-infrastructure-at-team8-3ff03cd86871) Team8 · Tel Aviv-Yafo, Israel · 2 weeks ago
- PE [Sr. AI Backend & Data Engineer](https://gurify.com/job/sr-ai-backend-data-engineer-at-pendo-3b7fe62caeba) Pendo · Herzliya, Israel · 3 weeks ago

```json
{"@context":"https://schema.org/","@type":"JobPosting","title":"Senior Solutions Engineer \u2013 Big Data \u0026 Data Infrastructure","description":"\u003Cp\u003EThis is a great opportunity to be part of one of the fastest-growing infrastructure companies in history, an organization that is in the center of the hurricane being created by the revolution in artificial intelligence.\u003C/p\u003E\u003Cp\u003E\u0026quot;VAST\u0026#39;s data management vision is the future of the market.\u0026quot;- Forbes\u003C/p\u003E\u003Cp\u003EVAST Data is the data platform company for the AI era. We are building the enterprise software infrastructure to capture, catalog, refine, enrich, and protect massive datasets and make them available for real-time data analysis and AI training and inference. Designed from the ground up to make AI simple to deploy and manage, VAST takes the cost and complexity out of deploying enterprise and AI infrastructure across data center, edge, and cloud.\u003C/p\u003E\u003Cp\u003EOur success has been built through intense innovation, a customer-first mentality and a team of fearless VASTronauts who leverage their skills \u0026amp; experiences to make real market impact. This is an opportunity to be a key contributor at a pivotal time in our company\u2019s growth and at a pivotal point in computing history.\u003C/p\u003E\u003Cp\u003EWe are seeking an experienced Solutions Data Engineer who possess both technical depth and strong interpersonal skills to partner with internal and external teams to develop scalable, flexible, and cutting-edge solutions. Solutions Engineers collaborate with operations and business development to help craft solutions to meet customer business problems.\u003C/p\u003E\u003Cp\u003EA Solutions Engineer works to balance various aspects of the project, from safety to design. Additionally, a Solutions Engineer researches advanced technology regarding best practices in the field and seek to find cost-effective solutions.\u003C/p\u003E\u003Ch3\u003EJob Description:\u003C/h3\u003E\u003Cp\u003EWe\u2019re looking for a Solutions Engineer with deep experience in Big Data technologies, real-time data pipelines, and scalable infrastructure\u2014someone who\u2019s been delivering critical systems under pressure, and knows what it takes to bring complex data architectures to life. This isn\u2019t just about checking boxes on tech stacks\u2014it\u2019s about solving real-world data problems, collaborating with smart people, and building robust, future-proof solutions.\u003C/p\u003E\u003Cp\u003EIn this role, you\u2019ll partner closely with engineering, product, and customers to design and deliver high-impact systems that move, transform, and serve data at scale. You\u2019ll help customers architect pipelines that are not only performant and cost-efficient but also easy to operate and evolve.\u003C/p\u003E\u003Cp\u003EWe want someone who\u2019s comfortable switching hats between low-level debugging, high-level architecture, and communicating clearly with stakeholders of all technical levels.\u003C/p\u003E\u003Ch3\u003EKey Responsibilities:\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EBuild distributed data pipelines using technologies like Kafka, Spark (batch \u0026amp; streaming), Python, Trino, Airflow, and S3-compatible data lakes\u2014designed for scale, modularity, and seamless integration across real-time and batch workloads.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EDesign, deploy, and troubleshoot hybrid cloud/on-prem environments using Terraform, Docker, Kubernetes, and CI/CD automation tools.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EImplement event-driven and serverless workflows with precise control over latency, throughput, and fault tolerance trade-offs.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ECreate technical guides, architecture docs, and demo pipelines to support onboarding, evangelize best practices, and accelerate adoption across engineering, product, and customer-facing teams.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EIntegrate data validation, observability tools, and governance directly into the pipeline lifecycle.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EOwn end-to-end platform lifecycle: ingestion \u2192 transformation \u2192 storage (Parquet/ORC on S3) \u2192 compute layer (Trino/Spark).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EBenchmark and tune storage backends (S3/NFS/SMB) and compute layers for throughput, latency, and scalability using production datasets.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EWork cross-functionally with R\u0026amp;D to push performance limits across interactive, streaming, and ML-ready analytics workloads.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EOperate and debug object store\u2013backed data lake infrastructure, enabling schema-on-read access, high-throughput ingestion, advanced searching strategies, and performance tuning for large-scale workloads.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003ERequired Skills \u0026amp; Experience:\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003E2\u20134 years in software / solution or infrastructure engineering, with 2\u20134 years focused on building / maintaining large-scale data pipelines / storage \u0026amp; database solutions.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EProficiency in Trino, Spark (Structured Streaming \u0026amp; batch) and solid working knowledge of Apache Kafka.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ECoding background in Python (must-have); familiarity with Bash and scripting tools is a plus.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EDeep understanding of data storage architectures including SQL, NoSQL, and HDFS.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ESolid grasp of DevOps practices, including containerization (Docker), orchestration (Kubernetes), and infrastructure provisioning (Terraform).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EExperience with distributed systems, stream processing, and event-driven architecture.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EHands-on familiarity with benchmarking and performance profiling for storage systems, databases, and analytics engines.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EExcellent communication skills\u2014you\u2019ll be expected to explain your thinking clearly, guide customer conversations, and collaborate across engineering and product teams.\u003C/li\u003E\u003C/ul\u003E","identifier":{"@type":"PropertyValue","name":"Gurify","value":"senior-solutions-engineer-big-data-data-infrastructure-at-vastdata-3d50a623835e"},"url":"https://gurify.com/job/senior-solutions-engineer-big-data-data-infrastructure-at-vastdata-3d50a623835e","datePosted":"2026-07-22","validThrough":"2026-09-21T23:59:59Z","hiringOrganization":{"@type":"Organization","name":"Vastdata","sameAs":"https://www.comeet.com/jobs/vastdata"},"directApply":false,"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressCountry":"IL","addressLocality":"Tel Aviv-Yafo"}}}
```

```json
{"@context":"https://schema.org/","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Jobs","item":"https://gurify.com/jobs"},{"@type":"ListItem","position":2,"name":"Israel","item":"https://gurify.com/jobs/israel"},{"@type":"ListItem","position":3,"name":"Senior Solutions Engineer \u2013 Big Data \u0026 Data Infrastructure","item":"https://gurify.com/job/senior-solutions-engineer-big-data-data-infrastructure-at-vastdata-3d50a623835e"}]}
```
