Job Description
We are seeking a talented Staff Engineer to design, develop, and deliver scalable software solutions with a focus on Generative AI and modern cloud technologies. The ideal candidate will have strong software engineering fundamentals and hands-on experience building AI-powered applications using Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and modern AI frameworks. You will collaborate with cross-functional teams to translate business requirements into reliable, secure, and high-performing solutions while contributing to technical design, architecture, and engineering best practices.
Responsibilities- Design, develop, and maintain enterprise-grade Generative AI applications.
- Build Retrieval-Augmented Generation (RAG) solutions using vector databases and enterprise knowledge sources.
- Develop AI assistants, copilots, and intelligent workflows using LLMs and AI orchestration frameworks.
- Integrate foundation models from providers such as OpenAI, Anthropic, Azure OpenAI, Google Gemini, and open-source models.
- Design and implement REST APIs and microservices to support AI applications.
- Collaborate with cross-functional teams to translate business requirements into technical solutions.
- Optimize AI application performance, scalability, latency, and operational costs.
- Implement prompt engineering, model evaluation, and AI guardrails to improve solution quality.
- Follow software engineering best practices, including code reviews, automated testing, CI/CD, and DevOps.
- Contribute to technical design discussions and architecture decisions.
- Mentor junior engineers and promote engineering best practices within the team.
- Stay current with emerging AI technologies and recommend improvements to existing solutions.
- Bachelor's or Master's degree in Computer Science, Engineering, Artificial Intelligence, or a related field.
- 8+ years of software engineering experience, including hands-on experience with Generative AI technologies.
- Strong proficiency in Python and SQL.
- Experience building applications using Large Language Models (LLMs).
- Hands-on experience with Retrieval-Augmented Generation (RAG) and vector databases.
- Experience with AI orchestration frameworks such as LangChain, LlamaIndex, or Semantic Kernel.
- Experience developing REST APIs and microservices.
- Experience with Docker, Kubernetes, and cloud platforms (AWS, Azure, or GCP).
- Strong understanding of software design principles, distributed systems, and cloud-native application development.
- Experience with Git, CI/CD pipelines, and modern software development practices.
- Strong analytical, problem-solving, and communication skills.
- Relocation Assistance Provided: Yes
No Referrers Available
There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.
