Role Overview: Lead the development of the application layer for enterprise GenAI solutions. Connect LLM backends to scalable frontends while managing API gateways and cloud deployments.
Key Responsibilities
- Application Architecture: Design scalable microservices that handle LLM requests, streaming responses (Server-Sent Events), and context management.
- Cloud & DevOps: Oversee the deployment of AI applications on AWS, Azure, or GCP.
- Integrate CI/CD pipelines for AI software components.
- Frontend & Backend Integration: Ensure seamless, low-latency integration between modern frontends (React/Next.js) and Python/FastAPI backends running AI models.
Required Skills & Qualifications
- Tech Stack: Python (FastAPI/Django), JavaScript/TypeScript (React, Node.js), Docker,
- Kubernetes, AWS/Azure AI services.
- Qualifications: Bachelor’s/Master’s in CS; 4–7 years in full-stack development with a strong recent focus on integrating AI/ML models into web apps.