SR. CLOUD INFRASTRUCTURE ENGINEER (AI & LLM PLATFORMS)
Q6 Cyber
📍 United States, US0💼 Full-time🕐 6/3/2026
Apply now →Start your Pro trial in 30 seconds (3 days free): you also get the AI match score with your resume.
Role overview
Q6 Cyber is hiring for the SR. CLOUD INFRASTRUCTURE ENGINEER (AI & LLM PLATFORMS) role in United States, US. It is full-time, Senior level, in the Tech sector. It was posted 6/3/2026.
On TalentyGo you can review this job and apply more effectively: Charlie prepares an ATS-optimized resume and a cover letter tailored to "SR. CLOUD INFRASTRUCTURE ENGINEER (AI & LLM PLATFORMS)" at Q6 Cyber in about a minute. Before you apply, you can also check how well your profile fits, with a match score based on skills, experience, location and seniority.
- Role
- SR. CLOUD INFRASTRUCTURE ENGINEER (AI & LLM PLATFORMS)
- Company
- Q6 Cyber
- Location
- United States, US
- Work mode
- On-site
- Employment
- Full-time
- Seniority
- Senior
- Sector
- Tech
- Posted
- 6/3/2026
Description
Job Description
We are seeking a specialized Infrastructure Engineer to bridge the gap between our large data repositories, Cloud Platform and the rapidly evolving world of Large Language Models (LLMs). You will be responsible for building the "plumbing" that allows our internal teams and external users to leverage AI effectively. This includes deploying Model Context Protocol (MCP) servers, building agentic execution environments, and scaling our internal Retrieval-Augmented Generation (RAG) architecture.
roles and responsibilities
Key Responsibilities
AI Architecture Guidance: Guide the architecture that will allow us to leverage AI tools with our large existing data stores and incoming streams of realtime intelligence.
Cross-Team Integration: Work closely with other infrastructure engineers and software development teams to integrate AI tools into existing systems.
MCP Ecosystem Management: Design, deploy, and maintain Model Context Protocol (MCP) servers to allow LLMs to securely interact with our internal databases, APIs, and external tooling.
Agentic Infrastructure: Build and orchestrate sandboxed, scalable environments (e.g., using Docker or specialized runtimes) where users can safely build and execute AI agents.
Internal RAG Platform: Develop and manage the infrastructure for our internal RAG (Retrieval-Augmented Generation) pipeline, including vector database management (e.g., Pinecone, Weaviate, or pgvector) and automated embedding pipelines.
Deployment & Scaling: Utilize Kubernetes (K8s) and Infrastructure as Code (Terraform/Pulumi) to deploy LLM-related tools, ensuring high availability and low latency for model inference and data retrieval.
Security & Governance: Implement strict guardrails for data privacy within LLM workflows, ensuring internal datasets remain secure while being accessible to authorized AI tools.
Required Qualifications
Required Qualifications:
5+ years of experience in DevOps, Platform Engineering, or SRE, with at least 1-2 years specifically focused on AI/ML infrastructure.
Proven track record of building production-grade RAG pipelines or LLM-integrated applications.
Thrives in "day zero" environments where the tools and protocols (like MCP) are evolving weekly.
Deep understanding of the security implications of LLMs (prompt injection, data leakage, and secure tool execution).
Experience working with substantial datasets (over 1bn objects, dozens or hundreds of TBs) and the challenges of leveraging AI tools with these data sets.
Bachelor's degree or equivalent in computer science or related field.
Required Technical Skills
Cloud & Orchestration: AWS/GCP/Azure, Kubernetes, Terraform, Helm.
AI Frameworks: LangChain, LlamaIndex, LangGraph.
Data & Vectors: Pinecone, Milvus, Qdrant, or pgvector; Apache Kafka/Pulsar; Elasticsearch/OpenSearch; traditional SQL RDBMS.
Languages: Python (Expert), TypeScript/Node.js (for MCP development), Go.
AI Protocols: Model Context Protocol (MCP), REST/gRPC.
The market for this role in United States
TalentyGo lists 5941 similar roles (383 in United States), 23% remote. Charlie ranks them against your CV, each with a clear score.
Similar jobs
Software Engineer, API Agents
📍 San Francisco · tech
Software Engineer, Agent Productivity
📍 San Francisco · Remote · tech
Frontend Engineer III, Moderation Enforcement
📍 Remote - United States · Remote · tech
Software Engineer, API Multimodal
📍 San Francisco · tech
Member of Technical Staff (iOS Engineer, Computer Growth)
📍 San Francisco · tech
Member of Technical Staff (Android Engineer, Computer Growth)
📍 San Francisco · Remote · tech
TalentyGo is an aggregator of job postings from public sources. Always verify information directly with the company. Applications go through the original company website; TalentyGo does not manage hiring processes.