Member of Technical Staff, Infrastructure
Technologies
Terraform · CDKTF · Kubernetes · Helm · Prometheus · Grafana · New Relic · ArgoCD · Flux · Python · Postgres
About LlamaIndex
Data indexing and RAG framework for enterprises building AI agents on unstructured data; LlamaCloud SaaS for document parsing and knowledge retrieval at scale.
Series A · 50–100 people
Job description
The full responsibilities and requirements are on the employer’s site.
Open application page ↗Source and classification
Internal deployment & tooling · Evidence for this classification:
Join us and help shape the future of AI by defining the narrative around document understanding. About the Role The Infra team at LlamaIndex owns the foundations that our product is built upon as well as many of the tools that enable engineers to develop, ship, and observe their code. We are responsible for designing, building, and scaling core infrastructure that powers a high-volume data platform for AI applications. We are looking for team members who love building enabling systems that empower our engineers and power our rapidly growing product. We’re looking for folks with experience managing cloud infrastructure, working through various stages of scale, and helping the broader Engineering team be more effective and productive. Some traits that are important to our company culture: customer-obsessed, collaborative, hard-working, and optimistic. And we’re looking for owners, so
More from the job description
Join us and help shape the future of AI by defining the narrative around document understanding. About the Role The Infra team at LlamaIndex owns the foundations that our product is built upon as well as many of the tools that enable engineers to develop, ship, and observe their code. We are responsible for designing, building, and scaling core infrastructure that powers a high-volume data platform for AI applications. We are looking for team members who love building enabling systems that empower our engineers and power our rapidly growing product. We’re looking for folks with experience managing cloud infrastructure, working through various stages of scale, and helping the broader Engineering team be more effective and productive. Some traits that are important to our company culture: customer-obsessed, collaborative, hard-working, and optimistic. And we’re looking for owners, so we hope you’ll help us expand this list. Responsibilities Collaborate with other engineering teams to build and maintain foundational systems that empower developers and support the company's rapid growth. Design and implement scalable infrastructure solutions for various deployment models, including SaaS, single-tenant, and private deployments. Manage and optimize cloud resources and Kubernetes clusters for cost-effectiveness and performance. Enable external customer deployment success throu [... source excerpt omitted ...] nd reliability. Ensure compliance with relevant regulations and implement robust security measures across different deployment environments. Build and operate infrastructure for production LLM applications, including model integrations, inference workloads, evaluation pipelines, observability, and the reliable execution of agentic workflows. Qualifications 8+ years of engineering experience. Worked on Platform or Infrastructure teams on significant projects involving infrastructure components (Terraform/CDKTF, Kubernetes, Helm, test infrastructure, release management, observability, etc.) Experience in optimizing cloud resource utilization. Proficient in tuning Kubernetes cl [... source excerpt omitted ...] ure as we grow. You can balance speed and pragmatism and build the appropriate solutions for each stage of the company’s growth. Hands-on proficiency with modern LLM tooling and production AI systems, including experience with model APIs, agent or RAG frameworks, evaluation and tracing tools, and the operational characteristics of LLM workloads. Preferred Qualifications Experience building out infrastructure from the ground up at a fast-growing startup. Experience with observability tools like Prometheus, Grafana, and New Relic. Experience with GitOps tools like ArgoCD and Flux for continuous deployment. Experience with security compliance and audits in cloud environments su
Employer postings · Data from · Sources