Job description
This is a remote position.
We are seeking a highly experienced AI Architect with deep expertise in Generative AI, Large Language Models (LLMs), and modern AI orchestration frameworks such asLangChainandLangGraph. The ideal candidate will have strong hands-on experience designing enterprise-grade AI solutions, building scalable RAG (Retrieval-Augmented Generation) pipelines, integrating MCP servers using Python, and developing production-ready AI applications leveraging modern LLM ecosystems.
Requirements
12–15+ years of overall IT experience with strong architecture background.
Extensive hands-on experience with Python development.
Strong expertise in LangChain and LangGraph frameworks.
Experience integrating MCP servers using Python.
Deep understanding of RAG architecture and LLM application development.
Hands-on experience with vector databases such as Pinecone, Weaviate, ChromaDB, or FAISS.
Experience working with OpenAI, Anthropic, Gemini, Llama, or other LLM ecosystems.
Strong understanding of prompt engineering and AI orchestration patterns.
Experience developing REST APIs and microservices.
Familiarity with Docker, Kubernetes, and cloud platforms such as AWS, Azure, or GCP.
Experience with AI monitoring, evaluation, and optimization techniques.
Strong knowledge of scalable distributed systems and enterprise architecture.
Originally posted on Himalayas
Who can apply
Eligible countries: United States. Accepted UTC offsets: UTC-10, UTC-9, UTC-8, UTC-7, UTC-6, UTC-5, UTC+14. Review the full description for employer-specific work authorization, residency and schedule requirements.