Maintain and optimize production LLM/VLM services and microservices (Flask/FastAPI), manage async queues, deploy and monitor Uvicorn/Gunicorn hosts, integrate LLM routing tools, design prompt strategies, build feedback/evaluation pipelines, expose secure REST APIs, and track token usage, latency, and errors to ensure SaaS-grade AI performance.
Role Overview
We are looking for an AI Engineer to maintain and enhance the AI-driven backbone of the Sootra platform. This role involves ensuring production stability of LLM/VLM pipelines, optimizing model interactions, maintaining APIs and queues, and building feedback loops that continuously improve AI outputs.
Responsibilities
- Maintain and optimize LLM- and VLM-powered services for content generation, compliance scoring, and campaign testing.
- Manage and scale Flask/FastAPI microservices, ensuring high uptime and low latency.
- Maintain Dramatiq queues for async AI workflows, campaign generation, and pipeline orchestration.
- Deploy, monitor, and debug Uvicorn/Gunicorn-based hosting in production environments.
- Integrate with OpenRouter and equivalent LLM routing tools to balance cost, latency, and quality.
- Design and refine prompt engineering strategies for reliability, context-awareness, and compliance.
- Build and maintain feedback pipelines for AI model evaluation (human-in-the-loop scoring, automated quality checks, reinforcement).
- Expose and maintain REST APIs for AI services, ensuring secure, versioned endpoints.
- Collaborate with backend/frontend teams to keep microservice architecture aligned and maintainable.
- Track token consumption, latency, and error rates to ensure production-grade performance.
Required Skills
- Programming: Strong in Python, with experience in production-grade codebases.
- Frameworks: Flask (for APIs), FastAPI (optional), Uvicorn/Gunicorn for async hosting.
- Queues/Workers: Dramatiq (or Celery/RQ equivalent) for background jobs.
- AI/ML: Hands-on with LLMs and VLMs, including prompt engineering, fine-tuning, and evaluation.
- AI Infrastructure: Familiar with OpenRouter or equivalent LLM/VLM routing & fallback tools.
- Architecture: Experience designing and maintaining microservice architectures.
- APIs: Strong experience with REST API design (auth, rate limiting, documentation).
- Production: Dockerized deployments, CI/CD pipelines, logging/monitoring, error handling.
- Feedback Loops: Building structured evaluation/feedback systems for AI model performance.
- Cloud: AWS/GCP experience preferred (deployment, monitoring, scaling).
Experience
- 3–5 years as an AI Engineer or Python Backend Engineer working with production systems.
- Prior work with SaaS platforms, LLM/VLM integrations, or AI-first products is highly valued.
Demonstrated ability to maintain AI pipelines in production, not just prototypes.
Similar Jobs
Artificial Intelligence • Consumer Web • Edtech • Enterprise Web • HR Tech • Social Impact • Generative AI
Build and deploy AI-powered solutions for enterprise and campus customers. Scope environments, prototype agentic AI applications, design secure multi-tenant and hybrid architectures, and harden successful prototypes into production. Own identity, encryption, networking, CI/CD, observability, and production support within customer environments. Collaborate with product, AI, data, and program teams to standardize repeatable solutions. The role is highly customer-facing and requires regular domestic and occasional international travel.
Top Skills:
Api GatewaysAWSAws PrivatelinkCi/CdDockerDuckdbEncryptionFastmcpIamJavaKafkaKey ManagementKubernetesLangchainLanggraphMcpMtlsObservabilityPgvectorPostgresPythonRagTypescriptVpc Peering
Cloud • Security • Software • Cybersecurity • Automation
Lead planning and delivery of large, cross-functional Enterprise Technology & AI programs. Translate strategy into roadmaps, align stakeholders, manage risks, define success metrics, improve governance, and communicate status to leadership while operating across IT, Finance, Legal, Product, and Business operations.
Top Skills:
AgileAICsmEnterprise ArchitectureEnterprise Data PlatformsGitlabInfrastructurePmpSafeScrumSecurity
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Oversee maintenance and reliability engineering for pharmaceutical manufacturing equipment and facility systems. Lead preventive, predictive, and breakdown maintenance, root cause investigations, CAPA, reliability programs, TPM, OEE improvements, equipment installation, qualification, and lifecycle management. Ensure cGMP, SOP, safety, and documentation compliance while collaborating with Production, QA, QC, and vendors.
Top Skills:
AutocadAutomation SystemsGranulation SystemsP&IdPackaging EquipmentProcess ControlsTablet Compression Machines
What you need to know about the Delhi Tech Scene
Delhi, India's capital city, is a place where tradition and progress co-exist. While Old Delhi is known for its rich history and bustling markets, New Delhi is defined by its modern architecture. It's clear the region places a strong emphasis on preserving its cultural heritage while embracing technological advancements, particularly in artificial intelligence, which plays a central role in shaping the city's tech landscape, fueled by investments in research and development.



