Deskripsi Pekerjaan
Informasi lengkap tentang posisi dan persyaratan
Ringkasan Yukerja
Lowongan AI Platform Engineer di AHEAD kami kurasi dari Himalayas (kategori Teknologi & IT). Posisi ini ditandai sebagai remote — pastikan timezone dan syarat lokasi kandidat di deskripsi resmi. Yukerja.com bukan pemberi kerja — lamaran diproses di situs sumber resmi.
Key Responsibilities
- Kubernetes for AI/ML: Architect and manage Kubernetes clusters tailored to AI/ML workloads.
- GPU Orchestration: Implement Run:ai and operators for GPU resource orchestration and workload scheduling.
- Automation & Pipelines: Develop and maintain Python-based automation scripts and ML pipelines; automate infrastructure provisioning with Terraform and configuration management with Ansible.
- Notebooks & Collaboration: Create and manage Jupyter Notebooks for experimentation and collaboration.
- NVIDIA Integration: Integrate and optimize NVIDIA Enterprise Suite components (CUDA, NeMo Framework, Triton, TensorRT, GPU drivers) for accelerated computing.
- MLOps Practices: Establish and maintain MLOps best practices for model lifecycle management, CI/CD, and monitoring (e.g., MLflow, Kubeflow).
- Collaboration: Work closely with data scientists and platform engineers to ensure efficient resource utilization and scalability across environments.
Required Skills & Experience
- Strong proficiency in Python and experience with ML frameworks (TensorFlow, PyTorch).
- Hands-on experience with Kubernetes and container orchestration.
- Familiarity with Run:ai or similar GPU scheduling platforms.
- Expertise in Terraform and Ansible for infrastructure automation.
- Experience with Jupyter Notebooks for ML development.
- Knowledge of NVIDIA Enterprise Suite (CUDA, NeMo Framework, Triton, GPU drivers).
- Solid understanding of MLOps principles and tools (e.g., MLflow, Kubeflow).
- Background in deploying and scaling AI workloads in cloud or hybrid environments.
Qualifications
- 4+ years in platform architecture or solutions architecture, with 2+ years focused on AI/ML workloads.
- Experience with high-performance computing (HPC) environments.
- Familiarity with distributed training and model optimization techniques.
- Certification in Kubernetes or cloud platforms (AWS, Azure, GCP).
Why AHEAD:
Through our daily work and internal groups like Moving Women AHEAD and RISE AHEAD, we value and benefit from diversity of people, ideas, experience, and everything in between.We fuel growth by stacking our office with top-notch technologies in a multi-million-dollar lab, by encouraging cross department training and development, sponsoring certifications and credentials for continued learning.India Employment Benefits include:
Comprehensive health insurance coverage for employees, with options to extend coverage to dependents Paid time off and company holidays, along with additional leave benefits as per policy Flexible work arrangements, supporting work-life balance Learning and development opportunities to support continuous growth and upskilling Employee wellness initiatives and programs focused on physical and mental well-being Retirement and statutory benefits in line with India regulations Inclusive and people-first culture, with a strong focus on collaboration and ownershipOriginally posted on Himalayas