Consultant – AI, LLM/SLM & Scalable ML Systems

Summary

Remote AI Consultant guiding a team to build and deploy scalable ML/LLM/SLM applications end-to-end, covering model training (RoPE, MoE, RLHF, vLLM), RAG pipelines, custom agents, multimodal AI, and production MLOps on Kubernetes/GPU clusters.

About the Role
We are looking for an AI Consultant to guide the team for building and deploying scalable AI applications and platforms. The consultant will provide technical leadership, best practices, and deployment strategies to help transform concepts into production-ready systems. This role involves advising, mentoring, and directing development in areas such as real-time communication, multi-modal AI, chat applications, custom agent creation, and scalable cloud deployment.

Responsibilities
· Lead design, training, and deployment of ML/LLM/SLM models (classification, recommendation, ranking, generative).

· Guide creation of custom LLMs/SLMs using RoPE, MTP, MoE, KV caching, vLLM, and RLHF (GRPO, PPO) for optimization.

· Build and scale RAG pipelines, chat applications, and custom task-specific agents (LangGraph, Agno, Smol-agents) with vector databases.

· Deliver multimodal systems: text↔speech, real-time streaming (LiveKit), Vision Transformers for image↔text, OCR.

· Provide direction on low-code automation (n8n, LangFlow, MCP servers) for rapid workflows.

· Ensure enterprise-grade deployment across cloud/local with MLOps, CI/CD, monitoring, cost optimization.

Mentor and upskill teams, driving best practices in architecture, scaling, and AI product delivery.

Required Skills and Technologies
· Proven experience with LLM/SLM training, fine-tuning, and inference acceleration (vLLM, quantization, caching).

· Strong skills in RAG, chatbots, custom agents, multimodal AI (speech, vision), and OCR.

· Hands-on with scalable deployments (Kubernetes, GPU clusters, hybrid cloud).

· Deep knowledge of MLOps, observability, and production AI lifecycle.


Experience Requirements
More than 3 years of relevant experience in machine learning, AI, or related fields.

Work Location & Setup
This position is fully remote, allowing you to work from anywhere, preferably in Pune, Maharashtra, India. You will be part of a collaborative virtual team environment.

Employment Type & Duration
- This is an hourly contract position with a commitment of 20 hours per week. The duration of the project is not specified.

Compensation & Benefits
- Compensation will be competitive and commensurate with experience. Specific salary details will be discussed during the interview process.

Success Criterion
· Production-ready AI applications deployed with scalability, reliability, and efficiency.

· Team enabled with frameworks, playbooks, and best practices for sustainable AI delivery.

· Faster transition from POC → production with measurable business impact.

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available