
PythonLLMsAPIsSystem DesignArchitectureMentoring
We are looking for a Staff AI Engineer to drive technical architecture, strategy, and production deployment for our next-generation AI and LLM platform services.
Tech stack: Python, LLM APIs (OpenAI/Anthropic), FastAPI, Vector Databases (Pinecone/Qdrant), Docker, PyTorch, LangChain/LlamaIndex
What you will work on:
- Architect enterprise-ready AI services leveraging Large Language Models, embeddings, and vector stores.
- Build high-throughput Python FastAPI microservices integrating complex prompt orchestration and model evaluations.
- Lead system design for RAG pipelines, enterprise knowledge ingestion, fine-tuning, and semantic search.
- Evaluate AI model latency, inference accuracy, token costs, and context optimization strategies.
What we are looking for:
- Strong track record of engineering and scaling AI/LLM applications into production setups.
- Advanced expertise in Python, generative AI architectures, vector databases, and API development.
- Proven technical leadership experience driving company-wide technical strategy and mentoring engineering teams.