About this opportunity
Namely lists this Senior Al Engineer - India opportunity in mountain view, California. Review the employer’s description below for duties, qualifications and application requirements.
Job description
Overview
We are seeking an enthusiastic Senior AI Engineer to join our team and contribute to the development of intelligent talent technology. This is a hands‑on opportunity for those early in their career to gain experience with AI and machine learning in a real‑world, enterprise SaaS environment. You will work alongside experienced engineers and learn how large language models, recommender systems, and other AI solutions enhance the workplace.
Responsibilities
Architect, optimize, and maintain high-performance, asynchronous orchestration layers (FastAPI) integrated with advanced inference servers (vLLM, Triton Inference Server) to host and serve open‑source models efficiently.
Build and maintain an advanced semantic caching layer and dynamic routing engine that evaluates inbound prompts and dispatches them to the most cost‑effective or lowest‑latency model (managed services vs. self‑hosted) based on intent and token volume.
Operationalize and maintain consistent deployments across three primary geographical regions, managing asymmetrical environment topologies (single‑region Dev environments vs. globally distributed production clusters).
Implement enterprise‑grade rate‑limiting, secure payload routing, and advanced telemetry dashboards using Prometheus and Grafana to track token consumption, error rates, and custom system metrics for internal departmental chargebacks.
Qualifications
Bachelor's or Master's degree in Computer Science or a related field.
6–8 years in backend platform engineering, systems architecture, or MLOps/LLMOps infrastructure design.
Expert‑level Python skills with deep experience in concurrent/asynchronous programming patterns (FastAPI, Asyncio).
Hands‑on experience with AWS Bedrock, prompt/model routing architectures, embedding generation, and hosting/serving open‑source models from Hugging Face.
Proficiency with core AWS services (EKS, IAM, VPC), containerization (Docker, Kubernetes).
#J-18808-Ljbffr
Worksite address
mountain view, CA, 94039, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.