Member of Technical Staff, Infrastructure

LlamaIndex
San Francisco, CA
HybridCareer-pivot friendly

Who this role is best for

Best suited to senior infrastructure engineers with experience in cloud and Kubernetes, working in AI and SaaS environments.

Best fit for

  • Senior infrastructure engineers with a background in cloud optimization and Kubernetes cluster management
    — “Proficient in tuning Kubernetes clusters and cloud resources for cost and performance efficiency
  • Candidates who have shaped engineering culture at fast-growing startups and can balance speed with reliability
    — “Willing to build LlamaIndex’s engineering culture as we grow
  • Individuals experienced in deploying and operating production AI systems, including LLMs and observability tools
    — “Build and operate infrastructure for production LLM applications, including model integrations, inference workloads, evaluation pipelines, observability, and the reliable execution of agentic workflows

Things to consider

  • Must have 8+ years of engineering experience, likely indicating a long-term career commitment
    — “8+ years of engineering experience
  • Requires a strong presence in San Francisco, given the hybrid-friendly culture based in the downtown office
    — “We offer a hybrid-friendly culture based out of our downtown San Francisco office

How to stand out

  • Highlight experience in multi-cloud deployments and cloud resource optimization on your resume
    — “Experience in optimizing cloud resource utilization
  • Showcase your hands-on work with LLM tooling and production AI systems in interviews
    — “Hands-on proficiency with modern LLM tooling and production AI systems
  • Demonstrate your understanding of observability tools like Prometheus and Grafana
    — “Experience with observability tools like Prometheus, Grafana, and New Relic
  • Mention your familiarity with GitOps tools such as ArgoCD and Flux in your application materials
    — “Experience with GitOps tools like ArgoCD and Flux for continuous deployment
Pace · Fast PacedCollaboration · HighAutonomy · HighDecision Impact · TeamLevel · Senior

Derived from job-description analysis by Serendipath's career intelligence engine.

What success looks like

  • infrastructure design
  • cloud resource optimization
  • release processes
  • observability tools
Typical background
infrastructure engineeringplatform engineering

Skills & requirements

Required

Cloud InfrastructureKubernetesObservabilitySecurityRelease ManagementContinuous Deployment

Preferred

GitopsSecurity ComplianceMulti-cloud Deployments

Stack & domain

TerraformCdktfKubernetesHelmTest InfrastructureRelease ManagementObservabilityPrometheusGrafanaNew RelicGitopsArgocdFluxSecurity ComplianceSOC 2PythonPostgreSQLMulti-cloud DeploymentsCustomer-obsessedCollaborativeHard-workingOptimisticOwnerSpeedPragmatismProblem-solvingTeamworkCommunicationCloud InfrastructureAI ApplicationsLLM ToolingProduction AI SystemsModel ApisAgent Or RAG FrameworksEvaluation And Tracing ToolsOperational Characteristics Of LLM Workloads

About the role

Original posting from LlamaIndex via Ashby

Join us and help shape the future of AI by defining the narrative around document understanding.

ABOUT THE ROLE

The Infra team at LlamaIndex owns the foundations that our product is built upon as well as many of the tools that enable engineers to develop, ship, and observe their code. We are responsible for designing, building, and scaling core infrastructure that powers a high-volume data platform for AI applications. We are looking for team members who love building enabling systems that empower our engineers and power our rapidly growing product.

We’re looking for folks with experience managing cloud infrastructure, working through various stages of scale, and helping the broader Engineering team be more effective and productive. Some traits that are important to our company culture: customer-obsessed, collaborative, hard-working, and optimistic. And we’re looking for owners, so we hope you’ll help us expand this list.

RESPONSIBILITIES

  • Collaborate with other engineering teams to build and maintain foundational systems that empower developers and support the company's rapid growth.
  • Design and implement scalable infrastructure solutions for various deployment models, including SaaS, single-tenant, and private deployments.
  • Manage and optimize cloud resources and Kubernetes clusters for cost-effectiveness and performance.
  • Enable external customer deployment success through maintaining clear infrastructure boundaries and principles.
  • Optimize and improve the release and deployment processes to enhance efficiency and reliability.
  • Ensure compliance with relevant regulations and implement robust security measures across different deployment environments.
  • Build and operate infrastructure for production LLM applications, including model integrations, inference workloads, evaluation pipelines, observability, and the reliable execution of agentic workflows.

QUALIFICATIONS

  • 8+ years of engineering experience.
  • Worked on Platform or Infrastructure teams on significant projects involving infrastructure components (Terraform/CDKTF, Kubernetes, Helm, test infrastructure, release management, observability, etc.)
  • Experience in optimizing cloud resource utilization.
  • Proficient in tuning Kubernetes clusters and cloud resources for cost and performance efficiency.
  • Willing to build LlamaIndex’s engineering culture as we grow.
  • You can balance speed and pragmatism and build the appropriate solutions for each stage of the company’s growth.
  • Hands-on proficiency with modern LLM tooling and production AI systems, including experience with model APIs, agent or RAG frameworks, evaluation and tracing tools, and the operational characteristics of LLM workloads.

PREFERRED QUALIFICATIONS

  • Experience building out infrastructure from the ground up at a fast-growing startup.
  • Experience with observability tools like Prometheus, Grafana, and New Relic.
  • Experience with GitOps tools like ArgoCD and Flux for continuous deployment.
  • Experience with security compliance and audits in cloud environments such as SOC2.
  • Familiar with Python, Postgres, multi-cloud deployments

LOCATION

We offer a hybrid-friendly culture based out of our downtown San Francisco office.

WHY JOIN US?

  • Impactful Mission: Work on innovative AI products that redefine how knowledge is accessed and utilized.
  • Collaborative Team: Join a team of passionate individuals committed to pushing the boundaries of technology.
  • Growth Opportunities: Be at the forefront of the AI revolution, with ample opportunities to grow alongside our scaling organization.

ADDITIONAL BENEFITS

  • Competitive base salary and equity compensation
  • Comprehensive medical/dental/vision coverage for you and your family
  • Unlimited paid time off policy
  • Daily catered lunch and snacks in the San Francisco office

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

LlamaIndex does not accept unsolicited agency resumes. Please do not forward resumes to our jobs alias, employees, or any other organization location. LlamaIndex is not responsible for any fees related to unsolicited resumes.

Source: LlamaIndex careers (Ashby)

Similar roles