Role Overview
We are looking for a motivated Generative AI Engineer to join our team and contribute to building cutting-edge AI/ML solutions. In this role, you will work on multimodal AI systems combining text, speech, and vision, and help develop realtime AI agents, avatar animations, and other next-generation AI applications. This is an excellent opportunity for early-career engineers to gain hands-on experience with advanced AI models and technologies while collaborating with research, product, and design teams.
Responsibilities
- Assist in designing, training, and deploying AI/ML models for speech, NLP, and avatar animation.
- Preprocess and prepare data from video, audio, and text sources.
- Integrate AI models and APIs into production pipelines and platforms.
- Conduct experiments to improve latency, accuracy, and scalability of AI solutions.
- Collaborate closely with research, product, and design teams to develop creative AI-driven features.
- Stay updated on the latest AI/ML research, frameworks, and tools.
Requirements
- Experience with AI/ML models for speech, NLP, and avatar animation.
- Proficiency in Python and basic ML/DL principles.
- Experience with PyTorch or similar deep learning frameworks.
- Knowledge of Transformers, GANs, or diffusion models.
- Familiarity with Git, Jupyter notebooks, and Linux development environments.
- Strong problem-solving skills and eagerness to learn quickly.
- Ability to debug and prototype AI systems effectively.
Good to Have Skills
- Audio processing frameworks (e.g., librosa, torchaudio).
- LLM frameworks – LangChain or Hugging Face Transformers.
- Computer vision tools – OpenCV, MediaPipe.
- Interest in multimodal AI, speech technologies, or real-time AI applications.
Skills
- Python
- PyTorch
- Transformers
- Git
- Linux