Job Description

Job Summary

We are seeking a highly skilled Python & Generative AI Engineer (6–8 years experience) to design, build, and optimize scalable, cloud-native AI applications. This role blends hands-on development, system design, and applied AI engineering, with a focus on Large Language Models (LLMs) and Small Language Models (SLMs).

The ideal candidate is a strong backend engineer with experience in distributed systems and AI-driven applications, capable of owning features end-to-end while contributing to architecture, performance optimization, and model fine-tuning workflows.

 

Job Description

Key Responsibilities

  • Design and develop scalable, cloud-native applications using Python
  • Build and maintain APIs and backend services with a strong focus on performance, reliability, and maintainability.
  • Performed automated data validation, missing value detection, duplicate analysis, and anomaly detection using Pandas and Polars.
  • Developed business rule engines to identify KPI deviations, performance gaps, and data quality issues.
  • Integrated Large Language Models (Azure OpenAI/Llama) to generate executive summaries, business insights, and recommendations.
  • Develop and integrate Generative AI solutions, including LLM-based applications and prompt engineering techniques.
  • Contribute to the design of distributed and event-driven systems with high availability.
  • Collaborate with architects and senior engineers to implement scalable and extensible system designs.
  • Implement CI/CD pipelines, automated testing, and DevOps best practices.
  • Leverage AI-assisted development tools (e.g., GitHub Copilot) to improve development efficiency.
  • Ensure code quality through testing, reviews, and adherence to engineering standards.
  • Monitor and optimize applications using logging, monitoring, and observability tools.
  • Work closely with cross-functional teams to deliver AI-powered business solutions.
  • SLM fine-tuning for edge or cost-efficient deployments