Senior Software Engineer – AI Infrastructure at Kraken | Torre

Senior Software Engineer – AI Infrastructure

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Full-time

Legal agreement: Employment

Provide your expected compensation while applying
location_on
Remote (for United Kingdom residents)
Remote (for Brazil residents)
Remote (for Canada residents)
Remote (for Cyprus residents)
Shared by
Emma of Torre.ai
4 months ago

Responsibilities


About the roleKraken is a mission-focused company rooted in crypto values. As a Krakenite, you’ll join the mission to accelerate the global adoption of crypto, so that everyone can achieve financial freedom and inclusion. Kraken is a fully remote company with Krakenites in 70+ countries.The teamThe AI Infrastructure team builds and operates the production systems that power intelligent agents at scale. This team sits at the foundation of the agent platform, ensuring that model inference, orchestration, and execution layers are reliable, observable, and performant under real-world load. Working closely with the Agent Systems team and broader infrastructure partners, this group owns the core primitives that enable agents to safely operate across internal systems. This is a deeply production-oriented team where engineers build in Rust and operate services where performance and failure modes matter.The opportunityDesign and build the infrastructure layer powering AI agent systems in productionDevelop high-performance Rust services that handle model inference, orchestration, and executionArchitect scalable systems capable of supporting millions of users and high request throughputBuild reliable ML infrastructure and MLOps patterns for model deployment, evaluation, and monitoringDefine guardrails, observability, and failure handling for agent-driven workflowsOptimize latency, throughput, and cost across inference and orchestration layersPartner closely with the Agent Systems team to translate experimental prototypes into hardened production systemsContribute to foundational infrastructure decisions in a high-scale, high-impact environmentSkills you should HODL5+ years of experience building and operating high-scale production systemsStrong proficiency in Rust and systems-level programmingDeep understanding of distributed systems, reliability engineering, and performance optimizationExperience operating services serving millions of users or high-throughput workloadsFamiliarity with ML infrastructure, model serving, or MLOps in production environmentsExperience designing observability, monitoring, and failure recovery systemsStrong collaboration skills working across infrastructure and applied engineering teamsHigh ownership mindset in high-stakes production environmentNice to havesExperience building infrastructure for agent-based or LLM-powered systemsBackground in high-performance networking, async systems, or low-latency architecturesExperience with container orchestration and cloud-native infrastructureFamiliarity with evaluation frameworks and model performance monitoring at scaleExperience working in fast-moving 0→1 or platform-building teamsHiring process notesUnless a specific application deadline is stated, applications are accepted on an ongoing basis.Applicants are permitted to redact or remove information on their resume that identifies age, date of birth, or dates of attendance at or graduation from an educational institution.Kraken considers qualified applicants with criminal histories, assessing candidates consistent with the San Francisco Fair Chance Ordinance.Kraken may ask candidates to complete job-related skills or work-style assessments as part of the hiring process.Equal opportunityAs an equal opportunity employer, Kraken doesn’t tolerate discrimination or harassment of any kind.