Role Overview
We are seeking an AI Engineer to design, develop, and deploy AI-driven applications utilizing large language models. You will build autonomous agentic systems, architect RAG pipelines, and manage production workloads on cloud infrastructure to deliver cutting-edge AI solutions.
Responsibilities
- Design and deploy applications using LLMs like GPT, Llama, and Claude.
- Integrate LangChain and LangGraph for robust agent workflows and tool orchestration.
- Develop function-calling mechanisms and manage APIs for seamless service integration.
- Write clean, scalable code in Python using Pydantic and async frameworks.
- Deploy and manage AI workloads on Google Cloud Platform (GCP) including Vertex AI.
- Architect and optimize Retrieval-Augmented Generation (RAG) pipelines with vector stores and embeddings.
- Build large-scale scraping and ingestion pipelines using Playwright and BeautifulSoup.
Requirements
- Bachelor’s or Master’s degree in CS, Engineering, Data Science, or a related field.
- 1-5 years of experience in AI/ML with strong exposure to LLMs and agent frameworks.
- Proficiency in Python and experience with LangChain and LangGraph.
- Hands-on experience with GCP for deploying production AI workloads.
- Deep understanding of ML fundamentals, embeddings, and vector databases.
- Proven experience with function calling, API integrations, and RAG pipelines.
Nice to Have
- Familiarity with multi-modal models and fine-tuning methods like LoRA.
- Knowledge of MLOps and model monitoring practices.
- Experience with containerization using Docker and Kubernetes.
- Experience with scraping frameworks like Playwright.
Skills
- Python
- LangChain
- GCP
- RAG
- LLMs