Hiring for a Global IT service provider, in "AI- Forward Deployed Engineers" based out of Chennai/Bangalore
Experience: 8+ Years, with Model Optimization, Fine-Tuning & Strategic AI
Are you someone who doesn’t just use AI models—but optimizes, fine-tunes, and pushes them to their limits?
Model Fine-Tuning: Implement PEFT (Parameter-Efficient Fine-Tuning), LoRA, and QLoRA to
adapt open-source models (Llama 3, Mistral) to specific client domains.
• Optimization & Quantization: Perform model quantization to reduce inference costs and
latency without sacrificing quality. Manage Dense Vectors and embedding optimizations.
• State-of-the-Art Exploration: Continuously research and implement the latest advancements
(e.g., State Space Models, Long-Context optimizations) into client deliverables.
• Strategic Consulting: Act as a trusted advisor to C-level client executives, defining the "Art of
the Possible" and guiding long-term AI roadmaps.
Technical Requirements:
• Deep Learning: PyTorch/TensorFlow, Transformers architecture internals, Attention
mechanisms.
• Model Ops: Serving custom models (vLLM, TGI), GPU memory management, Quantization
techniques (GGUF, AWQ).
• Advanced Data: Training data curation, synthetic data generation, RLHF concepts.
• Leadership: Ability to define the technical culture and set standards for the entire FDE
organization.