# Infrastructure Engineer (Linux)

[Fractile](https://gurify.com/jobs?q=Fractile) · London · Posted 2 weeks ago

[DevOps](https://gurify.com/jobs/devops)

[Apply on the original posting → (opens in a new tab)](https://job-boards.eu.greenhouse.io/fractile/jobs/4957200101)

## Job description

### Infrastructure Engineer (Linux)

### About Fractile

Fractile was founded in 2022 on the bet that, eventually, the world’s most capable AI systems would be limited in their impact by the time taken to produce useful outputs. We bet everything on the logical conclusion: that the only way to truly unlock this latent value, to make speed viable at scale, was to radically re-invent the hardware that we run our frontier AI models on. Ever since, we have been building chips and systems that tackle this problem: how to efficiently generate output at thousands of tokens per second, while handling the complexity and capacity challenges of operating large models at very long contexts.

The workloads that push to the limits of the current frontier are already transformational; it is the technical and economic limits on inference speed that are constraining progress. The defining work of the 21st century will be marked by the engine of inference delivering immense and diffuse chains of intellectual inquiry, in drug discovery, in software engineering, in materials discovery, in any field where progress is driven by deep reasoning and intelligence to resolve complex problems.

We are seeking an Infrastructure Engineer (Linux) to help build and maintain the compute, storage, and networking foundations that support our silicon development workloads, supporting an on-premise compute environment used for computationally-intensive engineering work, while growing your skills across Linux systems administration, HPC-style cluster management, networking, storage, and performance tuning. This is a great opportunity for someone with 2-3+ years of solid Linux experience who picks up new technical areas quickly, works methodically, and is comfortable operating independently, and who wants to develop deeper expertise in infrastructure, high-performance computing, and large-scale systems engineering.

### Key Responsibilities:

- Help maintain and troubleshoot a fleet of on-premise Linux (Rocky/RHEL-family) servers.

- Support the deployment and upkeep of on-premise compute infrastructure using infrastructure-as-code tooling (e.g. Ansible).

- Assist with monitoring and observability — setting up and maintaining tooling for resource utilisation, service health, and machine failures (e.g. Prometheus, node_exporter, Grafana, Zabbix).

- Help diagnose and resolve networking issues (DNS, VLANs, bonding, routing) and storage issues (network filesystems, capacity management, performance troubleshooting).

- Support day-to-day operation of a cluster compute/job scheduling environment (e.g. Slurm).

- Help with user and identity management tasks (e.g. FreeIPA/LDAP), including onboarding and access provisioning.

- Document infrastructure changes and contribute to runbooks and playbooks.

- Work with engineers across the organisation to understand and help resolve infrastructure-related bottlenecks.

### Requirements:

- 2-3+ years of solid, hands-on production Linux system administration experience.

- Good working knowledge of monitoring solutions (e.g. Prometheus, Grafana, Zabbix, or similar).

- Solid networking fundamentals — TCP/IP, DNS, VLANs, routing, basic troubleshooting.

- Good understanding of storage concepts — network filesystems, parallel/distributed filesystems, disk/volume management, basic performance troubleshooting.

- Comfortable working from the command line and with scripting (Bash, Python, or similar) to automate routine tasks.

- A methodical, curious approach to troubleshooting, with the ability to pick up new tools and technical areas quickly.

- Comfortable working independently and taking ownership of problems through to resolution.

### Desirable:

- Exposure to or knowledge of HPC (High Performance Computing) environments or cluster compute concepts.

- Familiarity with infrastructure-as-code tools (e.g. Ansible, Terraform).

- Linux certifications (e.g. RHCSA, LFCS, or similar).

- Experience working in a data centre environment.

- Knowledge of server hardware (e.g. installation, diagnostics, component replacement).

- Silicon industry background.

### How we work

- Ownership and execution: you will have full agency to drive your work forward

- Rapid iteration: we all work directly with top leadership to move from idea to hardware on ambitious timelines

- Full-stack engagement: hardware, software, silicon, and modelling teams all work closely together to create a product with generational impact

- Optimistic and pragmatic: we possess the will to win, and to do the hard work to get us there

- Team player mentality: the mission is bigger than any of us, and we have the curiosity and technical focus to see the best idea shipped, no matter who’s it is

### What we offer:

- Competitive salary: A competitive salary reflective of your experience and the specialist nature of the role.

- Equity & Ownership: meaningful equity so everyone shares in the value creation.

- Benefits: Private Medical, Dental and Vision, Contributory Pension, 25 Days holiday plus bank holidays and Life/Critical Illness Insurance.

• Diverse & fun office: we believe the hardest problems get solved by the broadest range of minds. We are committed to Equal Employment Opportunity through attracting and retaining a diverse team and building an inclusive environment.
Fractile is seeking to increase the clock speed of global progress, one chip at a time. We’ve recently raised $220M from investors including Founders Fund and Accel and our most important work lies ahead. Join us!

### Export controls

Our work involves technologies subject to UK, US and other international export control regulations. Certain roles may require additional eligibility checks to ensure compliance with applicable law. We'll be transparent about this throughout the hiring process.

**Live in Fractile’s hiring system.** Read from the company's own applicant tracking system, not reposted from a job board — so it's a real, open requisition rather than an ad that outlived the role.

We remove it as soon as it disappears at source.

## More jobs like this

- PA [Forward Deployed Infrastructure Engineer, New Grad](https://gurify.com/job/forward-deployed-infrastructure-engineer-new-grad-at-palantir-5c2cdf2ccee8) Palantir · London, United Kingdom · last week
- LD [Lead Devops Engineer - Waracle](https://gurify.com/job/lead-devops-engineer-waracle-3366c71a88fb) Dundee, United Kingdom · 3 days ago
- LO [Senior DevOps Engineer](https://gurify.com/job/senior-devops-engineer-at-localstack-08c24eb4c67f) Localstack · United Kingdom · last week
- EL [Lead DevOps Engineer](https://gurify.com/job/lead-devops-engineer-at-elliptic-d37f4210db40) Elliptic · London, United Kingdom · 4 weeks ago
- BL [DevOps Engineer (FTC)](https://gurify.com/job/devops-engineer-ftc-at-bluecrestcapitalmanagement-e0a576aea185) Bluecrestcapitalmanagement · London, United Kingdom · 2 weeks ago
- BL [AI DevOps Engineer](https://gurify.com/job/ai-devops-engineer-at-bluecrestcapitalmanagement-b926ee977b64) Bluecrestcapitalmanagement · London, United Kingdom · 2 weeks ago

```json
{"@context":"https://schema.org/","@type":"JobPosting","title":"Infrastructure Engineer (Linux)","description":"\u003Ch3\u003EInfrastructure Engineer (Linux)\u003C/h3\u003E\u003Ch3\u003EAbout Fractile\u003C/h3\u003E\u003Cp\u003EFractile was founded in 2022 on the bet that, eventually, the world\u2019s most capable AI systems would be limited in their impact by the time taken to produce useful outputs. We bet everything on the logical conclusion: that the only way to truly unlock this latent value, to make speed viable at scale, was to radically re-invent the hardware that we run our frontier AI models on. Ever since, we have been building chips and systems that tackle this problem: how to efficiently generate output at thousands of tokens per second, while handling the complexity and capacity challenges of operating large models at very long contexts.\u003C/p\u003E\u003Cp\u003EThe workloads that push to the limits of the current frontier are already transformational; it is the technical and economic limits on inference speed that are constraining progress. The defining work of the 21st century will be marked by the engine of inference delivering immense and diffuse chains of intellectual inquiry, in drug discovery, in software engineering, in materials discovery, in any field where progress is driven by deep reasoning and intelligence to resolve complex problems.\u003C/p\u003E\u003Cp\u003EWe are seeking an Infrastructure Engineer (Linux) to help build and maintain the compute, storage, and networking foundations that support our silicon development workloads, supporting an on-premise compute environment used for computationally-intensive engineering work, while growing your skills across Linux systems administration, HPC-style cluster management, networking, storage, and performance tuning. This is a great opportunity for someone with 2-3\u002B years of solid Linux experience who picks up new technical areas quickly, works methodically, and is comfortable operating independently, and who wants to develop deeper expertise in infrastructure, high-performance computing, and large-scale systems engineering.\u003C/p\u003E\u003Ch3\u003EKey Responsibilities:\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EHelp maintain and troubleshoot a fleet of on-premise Linux (Rocky/RHEL-family) servers.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ESupport the deployment and upkeep of on-premise compute infrastructure using infrastructure-as-code tooling (e.g. Ansible).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EAssist with monitoring and observability \u2014 setting up and maintaining tooling for resource utilisation, service health, and machine failures (e.g. Prometheus, node_exporter, Grafana, Zabbix).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EHelp diagnose and resolve networking issues (DNS, VLANs, bonding, routing) and storage issues (network filesystems, capacity management, performance troubleshooting).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ESupport day-to-day operation of a cluster compute/job scheduling environment (e.g. Slurm).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EHelp with user and identity management tasks (e.g. FreeIPA/LDAP), including onboarding and access provisioning.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EDocument infrastructure changes and contribute to runbooks and playbooks.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EWork with engineers across the organisation to understand and help resolve infrastructure-related bottlenecks.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003ERequirements:\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003E2-3\u002B years of solid, hands-on production Linux system administration experience.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EGood working knowledge of monitoring solutions (e.g. Prometheus, Grafana, Zabbix, or similar).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ESolid networking fundamentals \u2014 TCP/IP, DNS, VLANs, routing, basic troubleshooting.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EGood understanding of storage concepts \u2014 network filesystems, parallel/distributed filesystems, disk/volume management, basic performance troubleshooting.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EComfortable working from the command line and with scripting (Bash, Python, or similar) to automate routine tasks.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EA methodical, curious approach to troubleshooting, with the ability to pick up new tools and technical areas quickly.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EComfortable working independently and taking ownership of problems through to resolution.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EDesirable:\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EExposure to or knowledge of HPC (High Performance Computing) environments or cluster compute concepts.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EFamiliarity with infrastructure-as-code tools (e.g. Ansible, Terraform).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ELinux certifications (e.g. RHCSA, LFCS, or similar).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EExperience working in a data centre environment.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EKnowledge of server hardware (e.g. installation, diagnostics, component replacement).\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ESilicon industry background.\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EHow we work\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003EOwnership and execution: you will have full agency to drive your work forward\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ERapid iteration: we all work directly with top leadership to move from idea to hardware on ambitious timelines\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EFull-stack engagement: hardware, software, silicon, and modelling teams all work closely together to create a product with generational impact\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EOptimistic and pragmatic: we possess the will to win, and to do the hard work to get us there\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003ETeam player mentality: the mission is bigger than any of us, and we have the curiosity and technical focus to see the best idea shipped, no matter who\u2019s it is\u003C/li\u003E\u003C/ul\u003E\u003Ch3\u003EWhat we offer:\u003C/h3\u003E\u003Cul\u003E\u003Cli\u003ECompetitive salary: A competitive salary reflective of your experience and the specialist nature of the role.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EEquity \u0026amp; Ownership: meaningful equity so everyone shares in the value creation.\u003C/li\u003E\u003C/ul\u003E\u003Cul\u003E\u003Cli\u003EBenefits: Private Medical, Dental and Vision, Contributory Pension, 25 Days holiday plus bank holidays and Life/Critical Illness Insurance.\u003C/li\u003E\u003C/ul\u003E\u003Cp\u003E\u2022 Diverse \u0026amp; fun office: we believe the hardest problems get solved by the broadest range of minds. We are committed to Equal Employment Opportunity through attracting and retaining a diverse team and building an inclusive environment.\u003Cbr /\u003EFractile is seeking to increase the clock speed of global progress, one chip at a time. We\u2019ve recently raised $220M from investors including Founders Fund and Accel and our most important work lies ahead. Join us!\u003C/p\u003E\u003Ch3\u003EExport controls\u003C/h3\u003E\u003Cp\u003EOur work involves technologies subject to UK, US and other international export control regulations. Certain roles may require additional eligibility checks to ensure compliance with applicable law. We\u0026#39;ll be transparent about this throughout the hiring process.\u003C/p\u003E","identifier":{"@type":"PropertyValue","name":"Gurify","value":"infrastructure-engineer-linux-at-fractile-c25d5349864d"},"url":"https://gurify.com/job/infrastructure-engineer-linux-at-fractile-c25d5349864d","datePosted":"2026-08-25","validThrough":"2026-10-23T23:59:59Z","hiringOrganization":{"@type":"Organization","name":"Fractile","sameAs":"https://job-boards.eu.greenhouse.io/fractile"},"directApply":false,"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressCountry":"GB","addressLocality":"London"}}}
```

```json
{"@context":"https://schema.org/","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Jobs","item":"https://gurify.com/jobs"},{"@type":"ListItem","position":2,"name":"United Kingdom","item":"https://gurify.com/jobs/united-kingdom"},{"@type":"ListItem","position":3,"name":"Infrastructure Engineer (Linux)","item":"https://gurify.com/job/infrastructure-engineer-linux-at-fractile-c25d5349864d"}]}
```
