
What Blaxel does
Blaxel is the perpetual sandbox platform for AI agents that need to run code fast. Unlike traditional sandboxes that expire or suffer from cold starts, our sandboxes automatically scale to zero after 1s of inactivity and resume in under 25ms with full memory state preserved — allowing you to keep millions of secure sandboxes on standby indefinitely. We also co-locate your agent logic alongside these sandboxes to eliminate network overhead, delivering near-instant latency for end users.
6 open roles
What the role involves
We're looking for a world-class Forward Deployed Engineer to help strategic customers deploy Blaxel in real production environments. You will work directly with customer teams, founders, and Blaxel's Product & Engineering teams to turn ambiguous requirements into working systems: integrations, deployment plans, debugging workflows, runbooks, examples, and product fixes. This is a hands-on engineering role with direct customer ownership. Some days you'll be onsite at customer’s office mapping a customer's architecture and success criteria. Other days you'll be deep in code debugging an agent deployment, agent-LLM networking behavior, or sandbox runtime issues. The best person for this role can build trust in a room, then earn it again by shipping the solution. What you'll do Working side by side with our customers and interacting closely with Blaxel’s Product & Engineering teams, you’ll own the technical path from customer issues to successful deployment. Lead end-to-end technical discovery & implementation for strategic customer deployments. Engage directly with customer stakeholders (from engineering teams to executives) to understand their issues, architecture constraints and success criteria. Build on Blaxel platform APIs, SDKs, sandboxes, CI/CD systems, observability tools, as well as customer infrastructure. Ensure customer success and retention by improving their workloads reliability, performance, cost, and operational visibility. Bring customer learnings back into product and engineering with enough technical detail to shape roadmap decisions. Who you are Customer-obsessed builder: You enjoy working directly with users and shipping solutions that solve real problems while ensuring retained success for the Blaxel product. Strong engineer: You can write production code, navigate unfamiliar codebases quickly, and make pragmatic technical tradeoffs. Systems generalist: You are comfortable moving between APIs, cloud infrastructure, containers, networking, observability, and application logic. Fast in ambiguity: You can turn incomplete context into a concrete plan, then iterate as you learn. Excellent communicator: You can explain tradeoffs clearly, align stakeholders, and keep momentum during tense debugging sessions. High ownership: You don't stop at diagnosis. You drive problems through implementation, validation, documentation, and follow-up. Skills and experience Strong proficiency in TypeScript/Node or Python. Experience with Go, Rust, or systems-level engineering is a plus. Experience building and shipping production software, ideally in backend, infrastructure, cloud, developer platform, or AI product environments. Comfort with APIs, Docker, Kubernetes or similar orchestration, AWS/GCP/Azure, Linux, networking fundamentals, and CI/CD. Useful context includes agent frameworks, sandboxes, LLM applications, model-provider routing, eval workflows, vector databases, workflow orchestration, SSO/RBAC, security reviews, private networking, or compliance constraints. Prior forward-deployed, solutions, platform, customer engineering, founding engineer, or high-context consulting experience is helpful but not required. About Blaxel Blaxel is AWS for AI agents. We're building a new kind of cloud optimized specifically for the demands of agentic AI, powered by a purpose-built serverless compute engine delivering sub-25ms cold starts. We process billions of AI agent requests and power the infrastructure behind coding agents and background AI systems for top AI startups. Traditional clouds weren't built for agentic workloads. We are. We recently raised a $7.3M seed round led by First Round Capital.
What the role involves
We're looking for a world-class Forward Deployed Engineer to help strategic customers deploy Blaxel in real production environments. You will work directly with customer teams, founders, and Blaxel's Product & Engineering teams to turn ambiguous requirements into working systems: integrations, deployment plans, debugging workflows, runbooks, examples, and product fixes. This is a hands-on engineering role with direct customer ownership. Some days you'll be onsite at customer’s office mapping a customer's architecture and success criteria. Other days you'll be deep in code debugging an agent deployment, agent-LLM networking behavior, or sandbox runtime issues. The best person for this role can build trust in a room, then earn it again by shipping the solution. What you'll do Working side by side with our customers and interacting closely with Blaxel’s Product & Engineering teams, you’ll own the technical path from customer issues to successful deployment. Lead end-to-end technical discovery & implementation for strategic customer deployments. Engage directly with customer stakeholders (from engineering teams to executives) to understand their issues, architecture constraints and success criteria. Build on Blaxel platform APIs, SDKs, sandboxes, CI/CD systems, observability tools, as well as customer infrastructure. Ensure customer success and retention by improving their workloads reliability, performance, cost, and operational visibility. Bring customer learnings back into product and engineering with enough technical detail to shape roadmap decisions. Who you are Customer-obsessed builder: You enjoy working directly with users and shipping solutions that solve real problems while ensuring retained success for the Blaxel product. Strong engineer: You can write production code, navigate unfamiliar codebases quickly, and make pragmatic technical tradeoffs. Systems generalist: You are comfortable moving between APIs, cloud infrastructure, containers, networking, observability, and application logic. Fast in ambiguity: You can turn incomplete context into a concrete plan, then iterate as you learn. Excellent communicator: You can explain tradeoffs clearly, align stakeholders, and keep momentum during tense debugging sessions. High ownership: You don't stop at diagnosis. You drive problems through implementation, validation, documentation, and follow-up. Skills and experience Strong proficiency in TypeScript/Node or Python. Experience with Go, Rust, or systems-level engineering is a plus. Experience building and shipping production software, ideally in backend, infrastructure, cloud, developer platform, or AI product environments. Comfort with APIs, Docker, Kubernetes or similar orchestration, AWS/GCP/Azure, Linux, networking fundamentals, and CI/CD. Useful context includes agent frameworks, sandboxes, LLM applications, model-provider routing, eval workflows, vector databases, workflow orchestration, SSO/RBAC, security reviews, private networking, or compliance constraints. Prior forward-deployed, solutions, platform, customer engineering, founding engineer, or high-context consulting experience is helpful but not required. About Blaxel Blaxel is AWS for AI agents. We're building a new kind of cloud optimized specifically for the demands of agentic AI, powered by a purpose-built serverless compute engine delivering sub-25ms cold starts. We process billions of AI agent requests and power the infrastructure behind coding agents and background AI systems for top AI startups. Traditional clouds weren't built for agentic workloads. We are. We recently raised a $7.3M seed round led by First Round Capital.
What the role involves
The role We're looking for a world-class Founding Design Engineer to own product design end-to-end at Blaxel: from the first exploration to production code running in front of real users. You'll design the experience of a new category of cloud computing, then build it yourself, using AI to ship at a speed that wasn't possible even two years ago. This is a designer-who-ships role, not a handoff role. We need someone who thinks like a mini design founder, treats the product as their own, and closes the full loop: watch how people use Blaxel, design what should change, implement it in production, measure it, repeat. What you'll do Working with the cofounders and leveraging AI as much as possible, you'll own the Blaxel Console product experience from end to end. Design and ship. Take features from concept to polished UI to merged PR. Your designs land in production and get used by real customers, not archived in a Figma graveyard. Use AI as your implementation engine. Pair with AI coding tools to translate your designs into our front-end codebase, accurately and fast, and get them through review like any other engineer. Monitor how humans and AI agents actually use the platform. Watch sessions, dig into usage data, and proactively propose how the UX should evolve. You don't wait for a product owner to hand you a ticket. Close the loop. Ship improvements, measure their impact on activation and platform usage, double down on what works, kill what doesn't. Raise the design craft bar everywhere users touch Blaxel: console, onboarding, docs, emails, marketing site. Build the foundations: design system, component library, and interaction patterns that scale with the product. Who you are Designer first. Exceptional taste and craft are the bar, and the one thing we won't compromise on. Taste is the rarest skill in this role and the hardest to learn. Obsessive about detail. Spacing, motion, microcopy, empty states, edge cases. You notice the 1px misalignment and it bothers you until it's fixed. Technical enough to own a codebase. Solid front-end fundamentals (TypeScript, React, CSS), and you understand how a production codebase is structured, reviewed, and shipped. Deeply AI fluent. You are extremely curious about AI, and use it as a force multiplier on real design and engineering judgment, not a substitute for it. You're genuinely hungry to push the limit of what one person can ship. Product-minded. You think in user problems and KPIs, not just screens. Comfortable with analytics, session replays, and forming your own hypotheses about what to build next. Mini design founder. Strong ownership, high ambition, and comfort with the pace and ambiguity of an early-stage startup. You'd rather own the outcome than execute a spec. Why us? Design a new kind of product: Blaxel is used by humans and by AI agents directly. The playbook for designing interfaces that serve both has yet to be written. You'll write it, at the cutting edge of AI infrastructure. Fast growth: We're growing extremely fast, and a substantial proportion of our revenue already comes from AI agents directly using us. We have strong customer love and close to zero churn on the product. Direct impact & autonomy: You'll own the product experience on Blaxel Console end to end, with the full backing of the founders. Tier-1 backing: We're repeat founders & backed by the best in the business (First Round, YC), giving us the resources and network to build a generation-defining company.
What the role involves
The role We're looking for a world-class Founding Design Engineer to own product design end-to-end at Blaxel: from the first exploration to production code running in front of real users. You'll design the experience of a new category of cloud computing, then build it yourself, using AI to ship at a speed that wasn't possible even two years ago. This is a designer-who-ships role, not a handoff role. We need someone who thinks like a mini design founder, treats the product as their own, and closes the full loop: watch how people use Blaxel, design what should change, implement it in production, measure it, repeat. What you'll do Working with the cofounders and leveraging AI as much as possible, you'll own the Blaxel Console product experience from end to end. Design and ship. Take features from concept to polished UI to merged PR. Your designs land in production and get used by real customers, not archived in a Figma graveyard. Use AI as your implementation engine. Pair with AI coding tools to translate your designs into our front-end codebase, accurately and fast, and get them through review like any other engineer. Monitor how humans and AI agents actually use the platform. Watch sessions, dig into usage data, and proactively propose how the UX should evolve. You don't wait for a product owner to hand you a ticket. Close the loop. Ship improvements, measure their impact on activation and platform usage, double down on what works, kill what doesn't. Raise the design craft bar everywhere users touch Blaxel: console, onboarding, docs, emails, marketing site. Build the foundations: design system, component library, and interaction patterns that scale with the product. Who you are Designer first. Exceptional taste and craft are the bar, and the one thing we won't compromise on. Taste is the rarest skill in this role and the hardest to learn. Obsessive about detail. Spacing, motion, microcopy, empty states, edge cases. You notice the 1px misalignment and it bothers you until it's fixed. Technical enough to own a codebase. Solid front-end fundamentals (TypeScript, React, CSS), and you understand how a production codebase is structured, reviewed, and shipped. Deeply AI fluent. You are extremely curious about AI, and use it as a force multiplier on real design and engineering judgment, not a substitute for it. You're genuinely hungry to push the limit of what one person can ship. Product-minded. You think in user problems and KPIs, not just screens. Comfortable with analytics, session replays, and forming your own hypotheses about what to build next. Mini design founder. Strong ownership, high ambition, and comfort with the pace and ambiguity of an early-stage startup. You'd rather own the outcome than execute a spec. Why us? Design a new kind of product: Blaxel is used by humans and by AI agents directly. The playbook for designing interfaces that serve both has yet to be written. You'll write it, at the cutting edge of AI infrastructure. Fast growth: We're growing extremely fast, and a substantial proportion of our revenue already comes from AI agents directly using us. We have strong customer love and close to zero churn on the product. Direct impact & autonomy: You'll own the product experience on Blaxel Console end to end, with the full backing of the founders. Tier-1 backing: We're repeat founders & backed by the best in the business (First Round, YC), giving us the resources and network to build a generation-defining company.
What the role involves
The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You’ll be building and operating the core systems that power agentic AI at scale. Your mission: keep our ultra-low-latency, stateful, serverless compute engine rock-solid as we serve billions of agent requests for the most sophisticated AI teams in the world. This role is highly technical and execution-heavy. You’ll own our reliability posture end-to-end—observability, performance tuning, incident ops, infrastructure health, and the automation systems that keep everything running smoothly. We want you to design new reliability systems, push the boundaries of automation, and continuously evolve the platform to meet the demands of next-generation AI workloads. If you're a builder who thrives on owning critical infrastructure at scale, this role is for you. What you'll do Collaborating closely with the founders, the infra team, and the dev team—and leveraging AI wherever it creates leverage—you will architect and operate the systems that keep Blaxel fast, resilient, and secure. Architect, operate, and continuously improve the core infrastructure powering our 25ms cold-start compute engine. Build and evolve our observability stack (metrics, traces, logs), ensuring we detect issues before users do. Define, monitor, and drive SLOs/SLIs across key system surfaces to maintain world-class reliability. Lead incident response with rigor: root cause analysis, post-mortems, and driving systemic fixes. Design and implement self-healing, automated operational systems to eliminate toil and scale ops. Work across compute, networking, storage, and sandboxed execution layers to tune performance under extreme workloads. Build automation and tooling—often with AI agents—to streamline operations, debugging, capacity planning, and failure prediction. Stress-test and push our systems to the edge: load testing, chaos engineering, and performance benchmarking. Own security best practices at the infrastructure layer, from sandboxed compute to network isolation. Partner with platform engineers to ensure reliability is designed into new features from day one. Who you are Deeply technical by default: Fluent across systems, cloud, networking, and distributed computing. You love debugging real failures, not theoretical ones. AI-fluent operator: You understand how AI systems behave under scale, their unique resource patterns, and the infrastructure challenges of agentic frameworks. Builder at heart: You want to invent new reliability systems—not just maintain existing ones. You thrive in a zero-to-one infra environment. High-velocity execution: You have a strong bias for action and a track record of shipping reliable systems quickly with excellent judgment. Automation-first mindset: You hate repeated manual work and instinctively reach for automation or AI-driven ops to scale yourself. Calm under pressure: When incidents hit, you operate with clarity, precision, and ownership. Data-driven engineer: You measure everything—latency, tail behavior, resource efficiency, reliability trends—and let data guide your decisions. Required skills 3+ years in SRE, DevOps, or infrastructure engineering roles Strong proficiency in at least one programming language such as Go, Rust, or Python Hands-on experience with a major cloud provider (AWS, GCP) Solid knowledge of Linux systems, networking fundamentals, and distributed systems Experience with bare-metal servers and datacenter operations (PXE/iPXE provisioning, IPMI/BMC, RAID/NVMe, SR-IOV, high-throughput networking) Experience with Kubernetes or similar orchestrators Familiarity with observability stacks (Prometheus, Grafana, ELK, Datadog) Experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins) Strong debugging, problem-solving, and incident-management skills Preferred Exper
What the role involves
The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure platform. You’ll be building and operating the core systems that power agentic AI at scale. Your mission: keep our ultra-low-latency, stateful, serverless compute engine rock-solid as we serve billions of agent requests for the most sophisticated AI teams in the world. This role is highly technical and execution-heavy. You’ll own our reliability posture end-to-end—observability, performance tuning, incident ops, infrastructure health, and the automation systems that keep everything running smoothly. We want you to design new reliability systems, push the boundaries of automation, and continuously evolve the platform to meet the demands of next-generation AI workloads. If you're a builder who thrives on owning critical infrastructure at scale, this role is for you. What you'll do Collaborating closely with the founders, the infra team, and the dev team—and leveraging AI wherever it creates leverage—you will architect and operate the systems that keep Blaxel fast, resilient, and secure. Architect, operate, and continuously improve the core infrastructure powering our 25ms cold-start compute engine. Build and evolve our observability stack (metrics, traces, logs), ensuring we detect issues before users do. Define, monitor, and drive SLOs/SLIs across key system surfaces to maintain world-class reliability. Lead incident response with rigor: root cause analysis, post-mortems, and driving systemic fixes. Design and implement self-healing, automated operational systems to eliminate toil and scale ops. Work across compute, networking, storage, and sandboxed execution layers to tune performance under extreme workloads. Build automation and tooling—often with AI agents—to streamline operations, debugging, capacity planning, and failure prediction. Stress-test and push our systems to the edge: load testing, chaos engineering, and performance benchmarking. Own security best practices at the infrastructure layer, from sandboxed compute to network isolation. Partner with platform engineers to ensure reliability is designed into new features from day one. Who you are Deeply technical by default: Fluent across systems, cloud, networking, and distributed computing. You love debugging real failures, not theoretical ones. AI-fluent operator: You understand how AI systems behave under scale, their unique resource patterns, and the infrastructure challenges of agentic frameworks. Builder at heart: You want to invent new reliability systems—not just maintain existing ones. You thrive in a zero-to-one infra environment. High-velocity execution: You have a strong bias for action and a track record of shipping reliable systems quickly with excellent judgment. Automation-first mindset: You hate repeated manual work and instinctively reach for automation or AI-driven ops to scale yourself. Calm under pressure: When incidents hit, you operate with clarity, precision, and ownership. Data-driven engineer: You measure everything—latency, tail behavior, resource efficiency, reliability trends—and let data guide your decisions. Required skills 3+ years in SRE, DevOps, or infrastructure engineering roles Strong proficiency in at least one programming language such as Go, Rust, or Python Hands-on experience with a major cloud provider (AWS, GCP) Solid knowledge of Linux systems, networking fundamentals, and distributed systems Experience with bare-metal servers and datacenter operations (PXE/iPXE provisioning, IPMI/BMC, RAID/NVMe, SR-IOV, high-throughput networking) Experience with Kubernetes or similar orchestrators Familiarity with observability stacks (Prometheus, Grafana, ELK, Datadog) Experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins) Strong debugging, problem-solving, and incident-management skills Preferred Exper
Roles are as last read from the company’s own listings. Openings close without notice — check the date on the listing before you spend an evening on the application.
Check the company’s own careers page — linked at the top — before a job board. A role appears there first, sometimes weeks before it is syndicated anywhere else.
Questions and experiences
Nobody has asked anything about Blaxel yet. If you have interviewed here, what you know is worth more to the next person than anything on the rest of this page.
Company facts compiled from public sources and last refreshed 9 September 2026. Details change; treat the company’s own site as the authority.