Job Summary:
We are seeking a highly skilled Generative AI Engineer to design, develop, and deploy AI-powered applications using Large Language Models (LLMs). The ideal candidate should have experience with prompt engineering, Retrieval-Augmented Generation (RAG), vector databases, AI agents, and modern GenAI frameworks.
Key Responsibilities:
- Design and develop Generative AI applications using LLMs.
- Build AI chatbots, virtual assistants, and AI automation solutions.
- Develop RAG pipelines using vector databases.
- Create and optimize prompts for LLMs.
- Fine-tune and evaluate open-source and commercial language models.
- Integrate OpenAI, Anthropic, Gemini, or Hugging Face models into applications.
- Build AI agents using LangChain, LangGraph, CrewAI, or AutoGen.
- Develop REST APIs and deploy AI models to production.
- Monitor model performance, latency, and cost optimization.
- Collaborate with product, engineering, and data science teams.
Required Skills:
- Strong proficiency in Python.
- Experience with OpenAI API, Azure OpenAI, Gemini API, Claude API, or Hugging Face.
- Hands-on experience with LangChain, LangGraph, LlamaIndex, or similar frameworks.
- Knowledge of Prompt Engineering and RAG (Retrieval-Augmented Generation).
- Experience with Vector Databases such as Pinecone, ChromaDB, Weaviate, FAISS, or Milvus.
- Knowledge of Embeddings and semantic search.
- Experience with FastAPI, Flask, or REST APIs.
- Familiarity with Docker, Git, and cloud platforms (AWS, Azure, or GCP).
- Understanding of AI ethics, hallucination mitigation, and model evaluation.
Preferred Skills:
- Experience with AI Agents (CrewAI, AutoGen, LangGraph).
- Fine-tuning open-source LLMs (Llama, Mistral, Gemma, Qwen, etc.).
- Knowledge of MLOps and CI/CD pipelines.
- Experience with SQL, NoSQL databases, and cloud deployment.
- Familiarity with multimodal AI (text, image, audio).