Subscribe to email job alerts
Advertise your job vacancies
Prepaid job ad packages
| Job | Normal cost | Discount | Cost | Saving |
|---|---|---|---|---|
| 4 | R2,000 | 27% | R1,460 | R540 |
| 6 | R3,000 | 29% | R2,130 | R870 |
| 8 | R4,000 | 31% | R2,760 | R1,240 |
| 12 | R6,000 | 35% | R3,900 | R2,100 |
Recruitment news
AI can accelerate HR. But should it make the decisions?
Dereck Sigamoney
Discovery Health sponsors women’s health conference
Maroefah Smith

AI Engineer
| Location: | Johannesburg |
| Remote work: | Only remote work |
| Type: | Permanent |
| Reference: | #GZ61533 |
| Company: | E-Merge IT Recruitment |
You’ll have the opportunity to fine-tune models such as Llama, Mistral and Qwen, build self-hosted inference environments, optimise AI systems for performance and cost, and develop scalable machine learning solutions using modern cloud, containerisation and MLOps practices.
What we are looking for:
- 3+ years of experience building and shipping machine learning or AI-powered systems in production
- 7 plus years’ commercial development experience
- Strong Python skills, with hands-on experience using frameworks such as PyTorch, TensorFlow, or similar
- Practical experience with large language models - prompting, fine-tuning, RAG, or agent frameworks (e.g., LangChain, LlamaIndex, or custom implementations).
- Hands-on experience training and fine-tuning open-source models (e.g. Llama, Mistral, Qwen) full fine-tuning or parameter-efficient methods (LoRA/QLoRA) and deploying them on your own infrastructure rather than relying solely on hosted APIs.
- Experience standing up self-hosted inference serving on your own infra (e.g. vLLM, TGI, Triton, Ray Serve), including GPU provisioning, batching, and cost/latency optimization
- Solid understanding of ML fundamentals: model evaluation, overfitting, data leakage, and experimentation methodology
- Experience with vector databases and embedding-based retrieval (e.g. Pinecone, Weaviate, pgvector, FAISS)
- Comfort working with cloud infrastructure (AWS, GCP, or Azure) and containerized deployments (Docker, Kubernetes), including GPU-backed compute.
- Familiarity with MLOps practices model versioning, CI/CD for ML, monitoring, and rollback strategies for both hosted and self-hosted models
- Strong software engineering fundamentals: clean code, testing, code review, and API design
- Proficiency working with AI coding assistants (e.g. Claude Code, GitHub Copilot, Cursor) to accelerate development, while critically reviewing and validating AI-generated code
- Experience with model quantization, distillation, or optimization techniques (e.g. GPTQ, AWQ, ONNX) for efficient self-hosted inference
- Background in NLP, computer vision, or recommendation systems.
- Experience with streaming data pipelines (Kafka, Spark) or feature stores.
- Contributions to open-source ML/AI tooling
- Prior experience in a startup or fast-growth environment
Contact Garth on az.oc.egrem-e@zhtrag or call 011 463 3633 to discuss this and other opportunities.
Are you ready for a change of scenery? We are a specialist niche recruitment agency. We offer our candidates a wide range of opportunities, ensuring we successfully match the right technology professionals with the right roles.
Do you know a developer or technology specialist looking for a new opportunity? We pay cash for successful referrals!
Posted on 14 Sep 11:48, Closing date 13 Nov
Or apply with your Biz CV
Create your CV once, and thereafter you can apply to this ad and future job ads easily.






