Siellium Global Labs Private Limited
Remote, IN
0-2 years
AI Trainer and Architecture Intern (Founding Team)
Competitive
Stand out. Get hired.
80% of applicants fail the initial ATS screen. Check your compatibility before you apply and stand out from the crowd.
AI Inference & Engine Optimization Intern (AI Training / PyTorch/CUDA)
About Siellium Global Labs
Siellium Global Labs Private Limited is a startup scaling a zero-retention infrastructure layer across sovereign enterprise networks. We are engineering a local-first execution layer that utilizes a dynamic Mixture-of-Agents (MoA) orchestration framework.
The Role:
We are looking for a highly capable AI Inference & Optimization Intern to build and manage our local inference pipelines. This is a pure backend infrastructure role. You will not be building standard UI wrappers; you will be directly manipulating GPU memory, optimizing token throughput, and deploying a high-performance Mixture-of-Agents (MoA) execution layer.
Your mission is to squeeze maximum computational power out of our hardware. You will be responsible for pinning base models into VRAM and executing the mathematical compression needed to allow multiple domain-specific expert models to swap in and out seamlessly on a per-request basis.
Core Responsibilities
Take the next step in your career. Apply directly on the company's platform.
Apply for this role