Why This Role Stands Out
You'll have a significant impact by owning and scaling the cloud infrastructure that powers Langfuse's rapidly growing platform, processing over a billion trace events monthly. This remote role is ideal for a proactive Senior Cloud Infra Engineer eager to build world-class observability and directly collaborate with the ClickHouse team, fostering exceptional career growth.
Quick Overview
Job Description
Why Cloud Infrastructure at Langfuse
Your work will keep Langfuse running — everywhere.
Langfuse processes over a billion trace events per month. When a Fortune 50 company relies on Langfuse in production, they're relying on the infrastructure you operate. You'll own uptime, performance, and cost efficiency across our entire cloud footprint — and you'll make sure every self-hosted deployment runs just as smoothly.
You'll operate Langfuse Cloud on AWS ECS Fargate and ClickHouse Cloud, with Datadog as the observability backbone. You'll also own our public self-hosted infrastructure — including our Helm chart, Docker Compose setup, and everything in between — so that teams from startups to enterprises can run Langfuse on their own terms.
This isn't a "maintain what exists" role. We're scaling fast, and you'll be the person who makes sure the infrastructure grows ahead of demand — not behind it.
Langfuse is now part of ClickHouse, which means the team behind the database at the core of our stack is one channel away. Few infrastructure roles give you that kind of direct access to the people who build your most critical dependency.
You will grow at Langfuse by
Own Langfuse Cloud operations: You'll run our production environments on AWS ECS Fargate and ClickHouse Cloud. You'll manage deployments, autoscaling, capacity planning, and cost optimization — making sure we stay fast and affordable as traffic scales.
Build world-class observability: You'll own our Datadog setup end to end — dashboards, alerts, and SLOs. When something degrades, you'll ensure we know before our customers do. You'll build the monitoring culture that lets the whole team ship with confidence.
Make self-hosting effortless: Thousands of teams run Langfuse on their own infrastructure. You'll own and evolve our Helm chart, Docker Compose configuration, and deployment documentation. You'll turn "works on my machine" into "works on every machine" — from a single-node setup to a multi-region enterprise deployment.
Automate everything: CI/CD pipelines, infrastructure-as-code, automated scaling, zero-downtime deployments. You'll replace manual processes with automation that makes the team faster and the platform more reliable.
Scale for what's next: We're growing fast and new product directions — like complex long-running agent observability and real-time evaluation — push the infrastructure in new ways. You'll be thinking ahead about what breaks at 10x scale and building the foundation before we get there. 10x is always just one quarter away here at Langfuse.
Harden security and compliance: As more enterprises adopt Langfuse, you'll help ensure our cloud and self-hosted deployments meet the security and compliance bar that large organizations require.
What we're looking for
Strong infrastructure or SRE engineer who gets excited about running systems at scale and making them better every day
Experience operating production workloads on AWS (ECS/Fargate, networking, IAM, S3, etc.) or on comparable hyperscale vendors.
Comfortable with container orchestration — Kubernetes and/or ECS, Helm charts, Docker
Experience with infrastructure-as-code (Terraform, Pulumi, CloudFormation, or similar)
Strong monitoring and observability instincts — you've built dashboards and alerts that actually caught problems (Datadog experience is a plus)
You organize yourself. You have strong opinions about reliability, automation, and how to ship infrastructure changes safely
Interest in open source software and genuine enjoyment helping users debug their self-hosted deployments
Thrives in a small, accountable team where your output is visible and matters
CS or quantitative degree preferred
Bonus points:
Experience with ClickHouse Cloud or other managed analytical databases
Background in operating high-throughput event processing or observability infrastructure
Contributions to open source infrastructure tooling (Helm charts, Terraform modules, etc.)
Former founder
Perks
Flexible work environment - ClickHouse is a globally distributed company and remote-friendly. We currently operate in over 25 countries.
Healthcare - Employer contributions towards your healthcare.
Equity in the company - Every new team member who joins our company receives stock options.
Time off - Flexible time off in the US, generous entitlement in other countries.
A USD$500 Home office setup if you’re a remote employee.
Global Gatherings – We believe in the power of in-person connection and offer opportunities to engage with colleagues at company-wide offsites.
Culture - We All Shape It
As part of a rapidly scaling start-up, you will be instrumental in shaping our culture.
Are you interested in finding out more about our culture? Learn more about our values here. Check out our blog posts or follow us on LinkedIn to find out more about what’s happening at ClickHouse.
Equal Opportunity & Privacy
ClickHouse provides equal employment opportunities to all employees and applicants and prohibits discrimination and harassment of any type based on factors such as race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Please see here for our Privacy Statement.
Similar jobs
- RO
Cloud Engineer
NewAuto ApplyRobco
Munich🇩🇪Hybrid2 days agoGCPAWSMLOps+7Technology - FI
Senior Cloud-Infrastructure Engineer (w/m/d)
NewAuto Applyfinanzen.net GmbH
Remote (Karlsruhe, Berlin, München)🇩🇪Remote2 days agoDynamoDBMySQLSpring+10Technology - WR
Teamleiter IT Infrastructure & Cloud (m/w/d)
Auto ApplyWAREMA Renkhoff SE
Marktheidenfeld, BY🇩🇪Hybrid4 days agoTechnology - BB
Cloud Engineer – Azure Local & Enterprise Backup Solutions (w/m/d)
Auto ApplyBEW Berliner Energie und Wärme GmbH
Berlin, BE🇩🇪Hybrid5 days agoOracleSQLAzure+1Technology - D&
Cloud Engineer / Cloud Administrator (w/m/d)
Auto ApplyDrees & Sommer SE
Düsseldorf🇩🇪On-site5 days agoGCPAWSAnsible+4Technology - D&
Cloud Engineer / Cloud Administrator (w/m/d)
Auto ApplyDrees & Sommer SE
Berlin🇩🇪On-site5 days agoGCPAWSAnsible+4Technology - YG
Cloud Security Engineer
Auto ApplyYgo
Remote EU🇩🇪Remote1 week agoMFAOAuthSOC 2+10Technology - IC
Senior Network Engineer
Auto ApplyIceye
Berlin🇩🇪Hybrid1 week agoAWSAgileAnsible+3Technology - RG
Google Cloud Data Engineer (m/f/d)
Auto ApplyREPA GROUP
Bergkirchen, BY🇩🇪Hybrid1 week agoGCPSQLETL+2Technology - IC
Senior Cloud Network Engineer
Auto ApplyIceye
Berlin🇩🇪Hybrid5 weeks agoAWSAgileGit+4Technology - ES
Senior Infrastructure Engineer – AI Platform (f/m/x)
Auto ApplyEye Security
Berlin - hybrid🇩🇪Hybrid2 weeks agoAWSComplianceContinuous Improvement+5Technology - JI
Senior Infrastructure Engineer
Auto ApplyJimdo.com
Germany🇩🇪Remote1 week agoSwiftAWSArgoCD+7Technology