Why This Role Stands Out
Advance your career at the forefront of AI research by building scalable agent orchestration and impactful evaluation environments at Chakra Labs. This role is perfect for a versatile engineer with deep infrastructure knowledge who thrives on creating innovative solutions and shaping the future of AI development. Apply now to join a dynamic team pushing the boundaries of artificial intelligence.
Quick Overview
Job Description
About Us
Chakra Labs' mission is to encode human taste into intelligence. We build high-fidelity environments, evals, and datasets for frontier AI research, working with several of the top labs.
Our work sits at the frontier of post-training, agent environments, data quality, and research infrastructure. We care about building systems that make models better in ways that are measurable, useful, and hard to fake.
What You'd Work On
Agent orchestration at scale. Hundreds of agent runs at once, each with its own stateful environment. 100M tokens per minute across the fleet. You own the dispatch layer: SQS, concurrency control, failure handling.
Environment and task design. We need environments that feel real and scenarios that actually push agents to their limits. You'd figure out how to build new evaluations and design the tasks that test what matters, not just what's easy to measure.
The product around the platform. Infrastructure nobody can use isn't infrastructure. You'd build the surfaces our customers touch - dashboards for run inspection, tooling for experts to use, APIs that make the platform feel obvious - and ship them end to end.
New frontiers. The agent evaluation space is moving fast. You'd stay on that edge, supporting new environment modalities and shipping integrations with external orchestration frameworks.
About You
Generalist range, infra depth. You're a strong engineer across the stack - backend services, data pipelines, enough frontend to ship a real interface - with genuine depth in systems. You'd rather own a whole problem than a layer of one.
Container orchestration. You're comfortable running Kubernetes or similar in production. Auto-scaling, pod lifecycle, persistent storage, networking. You can figure out why something won't schedule and reason about resource contention.
Distributed systems. You've built or maintained message-driven architectures. SQS, Kafka, or similar. You know how to keep jobs moving when things back up, retry without duplicating, and fail without losing work.
LLM infrastructure. You've run LLM workloads at scale. Token instrumentation, rate limit handling, prompt caching, multi-provider routing. You've built the plumbing between models and external tools, and you know what it takes to keep it all running under load.
Experience. No hard rule. Ideally at least 3 years at this level, but less works if the above sounds like you.
What Makes This Different
It's infra, but the workload is AI agents. You're monitoring model behavior alongside pod health, debugging token throughput alongside network throughput.
Our customers are AI researchers and labs. You'd work directly with the people pushing the frontier of what agents can do, and build the infrastructure they run it on.
Ownership, not theater. You own whole systems, not tickets in a queue. One week you're shipping a new environment type, the next you're scaling the dispatch layer to handle 10x the throughput. You will ship your work to real customers, and build things that didn't exist a month ago.
The team. Our team is ex-Stripe, Snap, AWS, Microsoft, Airtable — you'll work with a small team who has years of shipping high-impact products over the last decade.
Cutting edge. You will get to touch the latest and greatest technologies across the data, AI, and infrastructure stack.
Similar jobs
- VF
Electronic Trading Engineer
NewAuto ApplyVirtu Financial
New York🇺🇸Hybrid8 hours agoSQLC++Java+2Engineering - QU
Software Engineer I
NewAuto ApplyQualtrics
Provo🇺🇸Hybrid7 hours agoMySQLNode.jsPHP+13Technology - PT
Staff Software Engineer, Applied AI
NewAuto ApplyPeregrine Technologies
New York🇺🇸$200k - $275k/yrHybrid9 hours agoDjangoAWSMachine Learning+10Technology - PT
Staff Software Engineer, Applied AI
NewAuto ApplyPeregrine Technologies
San Francisco🇺🇸$200k - $275k/yrHybrid9 hours agoDjangoAWSMachine Learning+10Technology - NE
Staff Software Engineer, Tech Lead - NetBox Delivery
NewAuto ApplyNetboxlabs
US East Coast Remote🇺🇸Remote8 hours agoDjangoRustAWS+12Technology - NI
Sr. Software Engineer II, FCM
NewAuto ApplyNinjaTrader
Chicago or Remote*🇺🇸$140k - $190k/yrRemote7 hours agoADPDerivativesGoogle Cloud+3Technology - NR
Software Engineer II - Query Gateway
NewAuto ApplyNew Relic
Portland🇺🇸$126k - $158k/yrHybrid8 hours agoDockerGCPAWS+8Technology - ME
Software Engineer (L2) - Platform Integrations
NewAuto ApplyMeridianlink
US Remote🇺🇸Remote10 hours agoDockerAPI GatewayAzure+5Technology - MI
Software Engineer
NewAuto ApplyMintmcp
San Mateo🇺🇸Hybrid10 hours agoMachine LearningSnowflakeCockroachDB+4Technology - KR
Senior Software Engineer - React Native - Consumer
NewAuto ApplyKraken.com
United States🇺🇸Remote9 hours agoSwiftKotlinReact+1Technology - HO
Principal Software Engineer, Database Platform
NewAuto ApplyHorizon3ai
United States🇺🇸$255k - $290k/yrRemote7 hours agoNeo4jSQLAWS+10Technology - HP
Senior/Staff Software Engineer, Connected Systems
NewAuto ApplyHeron Power
Scotts Valley🇺🇸$163k - $207k/yrOn-site9 hours agoRustEdge ComputingFleet Management+7Technology