Data & AI
GenAI / AI Engineer
Primary Technology Stack
PythonPyTorchLLMs (GPT/Claude)LangChainVector Databases (Pinecone/Chroma)FastAPI
About the Role
We are hiring two Generative AI Engineers to build intelligent features, orchestrate model pipelines, and implement production-grade AI features. You will design retrieval-augmented generation (RAG) structures, fine-tune models, and deploy high-performance endpoints.
Key Responsibilities
- Architect and optimize RAG pipelines utilizing vector databases and semantic indexing systems.
- Deploy and configure custom large language models (LLMs) and embeddings parameters.
- Build backend APIs (FastAPI/Python) to route agent parameters, manage sessions, and stream outputs.
- Track and mitigate model issues, minimizing errors, prompt injections, and token costs.
- Collaborate with product squads to identify, design, and validate AI enablement opportunities.
Required Qualifications
- 4+ years of software engineering experience, with 2+ years dedicated to AI/ML model integrations.
- Proficient in Python, including libraries like LangChain, LlamaIndex, PyTorch, and Hugging Face.
- Experience configuring and querying production vector databases (Pinecone, Chroma, pgvector).
- Practical experience managing LLM parameters, chunking logic, and semantic architectures.
- Masters or Bachelors in Computer Science, Data Science, or related engineering focus.
Apply for this Position
Role: GenAI / AI Engineer
Secured SSL Endpoint. Candidate documents handled securely.