Senior/Staff Software Engineer – Platform and Execution Model
Posted 1ds ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
Senior/Staff engineer building Trase OS, the distributed platform powering enterprise AI agents in regulated industries. Designing reliable execution, security, governance, and observability primitives.
Responsibilities:
- Build and own critical parts of Trase OS, the shared platform powering Trase deployments in regulated environments
- Design and implement clean, durable abstractions for the platform execution model
- Ensure correctness and determinism in workflow execution
- Translate evolving product requirements into coherent platform architecture
- Build reliable distributed systems across retries, restarts, partial failures, and concurrent execution
- Develop platform APIs and SDKs connecting workflows, agents, tools, and product surfaces; drive versioning and compatibility
- Guarantee correctness through idempotency, deterministic replays, compensating actions, and data integrity
- Engineer reliability at scale through concurrency controls, rate limits, backpressure, sharding/partitioning, and workload isolation
- Build security and governance into the core through RBAC/ABAC, policy enforcement, fine-grained audit, and lineage
- Deliver observability through distributed tracing, structured logs, metrics, evaluation hooks, and an explainable trail of agent actions
- Own quality through design reviews, test strategy, performance baselines, SLOs, incident response, and postmortems
- Mentor engineers and establish strong engineering patterns across the platform
- Staff-level candidates lead architecture across teams, establish engineering standards, and mentor other engineers
- Some travel to customers may be required
Requirements:
- 8+ years of experience building distributed/platform systems, including significant experience defining architecture across teams or domains
- 4+ years owning mission-critical runtimes or workflow/orchestration systems
- Deep expertise with durable execution, including state machines, event sourcing, saga/compensation, idempotency, and exactly/at-least-once semantics
- Proven track record with security and governance in production systems, including auth, RBAC, audit, and policy
- Hands-on experience with observability, including Grafana or equivalent and trace correlation across async boundaries
- Strong systems design across storage, queues, schedulers, and evented architectures; performance tuning under load
- Excellence in a modern language such as Go, Rust, Java, or TypeScript and cloud-native stacks including containers, CI/CD, and IaC
- Comfortable operating in regulated or high-assurance environments; bias toward correctness, clarity, and documentation
- Strong technical judgment and ability to influence design decisions within a team and across closely related engineering areas
- Ability to incorporate advanced LLM capabilities into system design and platform architecture decisions where appropriate
Benefits:
- Career track opportunity with potential for rapid advancement with strong performance as the firm grows
- 100% employer paid, comprehensive health care including medical, dental, and vision for you and your family
- Paid maternity and paternity for 14 weeks at employees' normal pay
- Unlimited PTO, with management approval
- Opportunities for professional development and continued learning
- Optional 401K, FSA, and equity incentives available
- Mental health benefits are available through Tara Mind
- Cost effective GLP-1 solutions available through Crux

















