
Machine Learning Engineer – AI
Posted 3 days ago

Posted 3 days ago
This is a fully remote position, open to applicants in India.
• Refine and train SLMs utilizing Hugging Face, TRL, and adapter techniques such as LoRA, QLoRA, and PEFT.
• Enhance models for inference through quantization, pruning, and knowledge distillation.
• Deploy models to edge devices, mobile platforms, and local servers while adhering to strict latency requirements.
• Construct comprehensive end-to-end MLOps pipelines from data ingestion to deployment.
• Oversee model accuracy, latency, and hardware usage in a production environment.
• Assess model quality via benchmarking frameworks and custom evaluation suites.
• Proven experience in training and fine-tuning SLMs with Hugging Face.
• Familiarity with adapters and adapter techniques, including LoRA, QLoRA, and PEFT.
• Proficiency in implementing quantization, pruning, knowledge distillation, and model optimization strategies.
• Experience in deploying models to edge devices, mobile applications, and local servers.
• Capability to construct end-to-end MLOps pipelines from data ingestion through to deployment.
• Competence in monitoring model accuracy, latency, and CPU/GPU performance during production.
• Preferred experience in deployment within edge or mobile environments.
• Familiarity with ONNX export and cross-platform inference is a plus.
• Knowledge of MLOps tools, including experiment tracking, model registries, and CI/CD processes for ML, is preferred.
• Commitment to equal employment opportunities and diversity and inclusion initiatives.
• Access to Global Egnyte Employee Communities (EECs) that promote representation and inclusion.
• Chance to contribute to a secure multi-cloud content security and governance platform.
Get handpicked remote jobs straight to your inbox weekly.