AI Architect
Posted 13hrs ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
AI Architect scaling AI pipeline infrastructure for The Information, a technology and business journalism publisher. Designing reliable LLM workflows, observability, and high-fan-out platform capabilities.
Responsibilities:
- Maintain and operate all existing AI pipelines on the platform
- Migrate remaining workflows from the legacy orchestration framework to the primary platform
- Own on-call response for AI pipeline failures, including failed-job triage and retries
- Manage LLM provider relationships, API key rotation, cost tracking, and model upgrades
- Design and implement new high-fan-out AI pipelines for future horizontal workflows
- Establish conventions for onboarding new pipelines, including registry entries, job subclassing, batch fan-out, and automated reporting
- Drive architectural decisions around data storage, queue partitioning, concurrency throttling, and cost controls
- Evaluate and integrate new LLM providers and embedding models while maintaining backward compatibility with existing vector data
- Build observability and operational tooling, including custom reporting, cost tracking, and alerting
- Partner with product, editorial/content, and growth teams to translate requirements into pipeline designs
- Integrate AI pipelines with the broader technical stack
- Mentor engineers on AI pipeline patterns, prompt engineering, and platform architecture
- Own the technical proposal process for new pipelines and major infrastructure changes
Requirements:
- 5+ years building production applications in a modern web framework such as Rails, Django, or similar
- Strong experience with ORM usage, API-only application design, and background job processing
- Proven experience designing idempotent, retryable, fan-out job pipelines, including batching, concurrency controls, dead-letter handling, and queue partitioning
- Hands-on experience with multiple LLM providers, structured output, embeddings, prompt engineering, and cost tracking
- Experience with a vector database for similarity matching at scale, including batched queries, namespace management, and embedding model migration
- Experience running internal services on a cloud PaaS or equivalent, with relational databases, caching/queueing infrastructure, and observability tooling
- Track record of authoring technical design documents, making build-vs-buy decisions, and designing extensible platforms
- Must be eligible to live and work in the United States
- Preferred: 5+ years of Ruby/Ruby on Rails experience
- Preferred: Experience migrating workflows from a Python-based orchestration framework to a Ruby-based framework
- Preferred: Familiarity with agent orchestration and LLM tracing/observability ecosystems
- Preferred: Experience with AI alerting or notification systems
- Preferred: Background in media/publishing AI applications
- Preferred: Experience with state machine libraries for managing job lifecycles
- Preferred: Experience building retrieval-augmented generation (RAG) pipelines
Benefits:
- Company-paid medical, dental, and vision coverage for employees and their dependents
- Medical coverage that includes fertility care and $0 copays for in-office mental health visits with in-network providers
- Paid parental leave
- Generous paid time off (PTO) that increases with tenure
- 401(k) plan with employer matching contributions
- Flexible Spending Accounts (FSAs) for healthcare and dependent care expenses
- Fitness and wellness stipend
- Monthly cell phone reimbursement
- Company-sponsored lunches in the office every Monday
- Commuter benefits
- Supportive, inclusive, and diverse work environment with a zero-tolerance policy for harassment
- Bonus














