Role Overview
As an AI Engineering Intern at Novetum, you'll work directly with our core engineering team to design, build, and evaluate AI-powered systems that solve real problems at scale. This is a hands-on, high-autonomy role with meaningful responsibility from day one, focused on building real-world AI systems that create impact.
Responsibilities
- Design and develop AI agents and coding agents using leading LLMs, integrating tool use, memory, and multi-step reasoning.
- Build and optimize Retrieval-Augmented Generation (RAG) pipelines — from chunking and embeddings to vector stores and reranking.
- Implement Model Context Protocol (MCP) integrations and custom tool harnesses to extend LLM capabilities into real-world systems.
- Craft and systematically refine prompts (zero-shot, few-shot, chain-of-thought).
- Build AI workflows that orchestrate multiple models, tools, and data sources into end-to-end applications.
- Construct evaluation (eval) frameworks to benchmark model accuracy, reliability, and regression.
- Instrument AI systems with observability tooling — traces, token usage, latency, failure modes, and cost tracking.
- Contribute to generative AI applications across text, code, and multimodal contexts.
Requirements
- Currently pursuing or recently completed a B.Tech / M.Tech / MS / PhD in Computer Science, AI/ML, or a related field.
- Strong Python programming skills, with clean, readable, production-aware code.
- Familiarity with at least one LLM orchestration framework (LangChain, LlamaIndex, CrewAI, or Haystack).
- Exposure to vector databases (Pinecone, Weaviate, Chroma, or pgvector) and embedding models is a plus.
- Strong written communication skills.
- Ability to work full-time on-site in Hyderabad, India.
Skills
- Python
- Large Language Models (LLMs)
- Retrieval-Augmented Generation (RAG)
- LangChain
- Vector Databases