Senior / Staff Backend Engineer - AI Agents Platform
Engineering · Posted 8 months ago
We are looking for a Senior or Staff Backend Engineer with 6+ years of experience to join our AI Agents Platform team to help build our large-scale, highly reliable platform for vertical AI agents. We are seeking a product-minded engineer who is passionate about building scalable, high-performance systems and is excite...
- Way of working
- Hybrid
- Location
- Palo Alto, CA
- Pay range
- $225,000 to $300,000
- Level
- Staff
- Experience
- 6+ years
- Type
- Full time
- Visa sponsorship
- Not offered for this role
- The company
- Software Development · 20 to 50 people
Skills that matter here
What you would be doing
- Own core backend services: Build and operate the agent runtime services (tool execution, state management, orchestration, retries, idempotency, rate limiting).
- Design robust APIs and contracts: Develop stable, well-versioned APIs (internal + external), including webhook/event interfaces and integration adapters.
- Model complex domain data: Design schemas and data access patterns that support agent memory/state, workflow history, audit trails, permissions, and multi-tenant isolation.
- Integrations at scale: Implement and maintain high-quality integrations (OAuth, webhooks, sync engines, connectors) with strong observability and failure handling.
- Platform performance and cost control: Optimize throughput/latency, queue backlogs, caching, storage, and inference/tool-call costs; build guardrails for runaway tasks.
- Reliability engineering: Define and improve SLIs/SLOs, implement end-to-end tracing, build safety rails (timeouts, circuit breakers, budgets), and run incident response.
- Raise engineering standards: Drive code quality, testing strategy, on-call hygiene, runbooks, postmortems, and mentoring.
The full description
We are looking for a Senior or Staff Backend Engineer with 6+ years of experience to join our AI Agents Platform team to help build our large-scale, highly reliable platform for vertical AI agents. We are seeking a product-minded engineer who is passionate about building scalable, high-performance systems and is excited to contribute to a platform that is revolutionizing an industry by deploying autonomous AI agents.
What you will do:
- Own core backend services: Build and operate the agent runtime services (tool execution, state management, orchestration, retries, idempotency, rate limiting).
- Design robust APIs and contracts: Develop stable, well-versioned APIs (internal + external), including webhook/event interfaces and integration adapters.
- Model complex domain data: Design schemas and data access patterns that support agent memory/state, workflow history, audit trails, permissions, and multi-tenant isolation.
- Integrations at scale: Implement and maintain high-quality integrations (OAuth, webhooks, sync engines, connectors) with strong observability and failure handling.
- Platform performance and cost control: Optimize throughput/latency, queue backlogs, caching, storage, and inference/tool-call costs; build guardrails for runaway tasks.
- Reliability engineering: Define and improve SLIs/SLOs, implement end-to-end tracing, build safety rails (timeouts, circuit breakers, budgets), and run incident response.
- Raise engineering standards: Drive code quality, testing strategy, on-call hygiene, runbooks, postmortems, and mentoring.
Interested in this one?
There is no apply button here on purpose. Tell us about yourself, we book a short call, and if this role fits we walk you through the company and ask before anything is sent. Always free for you.
Tell us about you