AI Architect

Posted 13hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

AI Architect scaling AI pipeline infrastructure for The Information, a technology and business journalism publisher. Designing reliable LLM workflows, observability, and high-fan-out platform capabilities.

Responsibilities:

  • Maintain and operate all existing AI pipelines on the platform
  • Migrate remaining workflows from the legacy orchestration framework to the primary platform
  • Own on-call response for AI pipeline failures, including failed-job triage and retries
  • Manage LLM provider relationships, API key rotation, cost tracking, and model upgrades
  • Design and implement new high-fan-out AI pipelines for future horizontal workflows
  • Establish conventions for onboarding new pipelines, including registry entries, job subclassing, batch fan-out, and automated reporting
  • Drive architectural decisions around data storage, queue partitioning, concurrency throttling, and cost controls
  • Evaluate and integrate new LLM providers and embedding models while maintaining backward compatibility with existing vector data
  • Build observability and operational tooling, including custom reporting, cost tracking, and alerting
  • Partner with product, editorial/content, and growth teams to translate requirements into pipeline designs
  • Integrate AI pipelines with the broader technical stack
  • Mentor engineers on AI pipeline patterns, prompt engineering, and platform architecture
  • Own the technical proposal process for new pipelines and major infrastructure changes

Requirements:

  • 5+ years building production applications in a modern web framework such as Rails, Django, or similar
  • Strong experience with ORM usage, API-only application design, and background job processing
  • Proven experience designing idempotent, retryable, fan-out job pipelines, including batching, concurrency controls, dead-letter handling, and queue partitioning
  • Hands-on experience with multiple LLM providers, structured output, embeddings, prompt engineering, and cost tracking
  • Experience with a vector database for similarity matching at scale, including batched queries, namespace management, and embedding model migration
  • Experience running internal services on a cloud PaaS or equivalent, with relational databases, caching/queueing infrastructure, and observability tooling
  • Track record of authoring technical design documents, making build-vs-buy decisions, and designing extensible platforms
  • Must be eligible to live and work in the United States
  • Preferred: 5+ years of Ruby/Ruby on Rails experience
  • Preferred: Experience migrating workflows from a Python-based orchestration framework to a Ruby-based framework
  • Preferred: Familiarity with agent orchestration and LLM tracing/observability ecosystems
  • Preferred: Experience with AI alerting or notification systems
  • Preferred: Background in media/publishing AI applications
  • Preferred: Experience with state machine libraries for managing job lifecycles
  • Preferred: Experience building retrieval-augmented generation (RAG) pipelines

Benefits:

  • Company-paid medical, dental, and vision coverage for employees and their dependents
  • Medical coverage that includes fertility care and $0 copays for in-office mental health visits with in-network providers
  • Paid parental leave
  • Generous paid time off (PTO) that increases with tenure
  • 401(k) plan with employer matching contributions
  • Flexible Spending Accounts (FSAs) for healthcare and dependent care expenses
  • Fitness and wellness stipend
  • Monthly cell phone reimbursement
  • Company-sponsored lunches in the office every Monday
  • Commuter benefits
  • Supportive, inclusive, and diverse work environment with a zero-tolerance policy for harassment
  • Bonus