About this role
Role Overview: Lead the development of the application layer for enterprise GenAI solutions. Connect LLM backends to scalable frontends while managing API gateways and cloud deployments.
Key Responsibilities
• Application Architecture: Design scalable microservices that handle LLM requests, streaming responses (Server-Sent Events), and context management. • Cloud & DevOps: Oversee the deployment of AI applications on AWS, Azure, or GCP. • Integrate CI/CD pipelines for AI software components. • Frontend & Backend Integration: Ensure seamless, low-latency integration between modern frontends (React/Next.js) and Python/FastAPI backends running AI models.
Required Skills & Qualifications
• Tech Stack: Python (FastAPI/Django), JavaScript/TypeScript (React, Node.js), Docker, • Kubernetes, AWS/Azure AI services. • Qualifications: Bachelor’s/Master’s in CS; 4–7 years in full-stack development with a strong recent focus on integrating AI/ML models into web apps.