Job description for AI Model Engineer at Alpha Red Pro Sdn Bhd
Job Highlights
- Build the next generation of Enterprise AI capabilities.
- Work with cutting-edge open-source AI technologies.
- Develop and fine-tune AI models instead of simply consuming AI APIs.
- Build AI infrastructure that powers multiple enterprise products.
- Grow your career in one of the fastest evolving technology fields.
Job Description
We are looking for a passionate AI Model Engineer to join our growing AI Engineering team. Unlike many organisations that simply integrate third-party AI services, Alpha Red Pro is investing in building our own AI platform and proprietary AI capabilities. Our vision is to develop, fine-tune, evaluate and deploy AI models that become reusable building blocks across multiple enterprise solutions in the Travel, Financial Services and Enterprise domains.
This role focuses on building the AI models and infrastructure that power our products. Rather than primarily consuming cloud-hosted AI APIs, you will be responsible for developing internal AI capabilities that can be adopted by software engineering teams throughout the organisation. If you enjoy experimenting with open-source models, understanding how they work, improving their performance and pushing the boundaries of enterprise AI, we would love to hear from you.
Key Responsibilities
- Deploy and manage open-source Large Language Models (LLMs).
- Evaluate, benchmark and compare AI models for different business use cases.
- Fine-tune domain-specific AI models using company datasets.
- Optimise model inference for performance, scalability and cost.
- Develop reusable AI services and model APIs for internal engineering teams.
- Design model evaluation and benchmarking pipelines.
- Research emerging AI technologies and recommend practical adoption.
- Collaborate with software engineers to integrate internally hosted AI models into enterprise applications.
- Continuously improve AI model quality through experimentation and iterative development.
- Document AI architectures, experiments and technical findings.
Job Requirements
- 2–5 years of software engineering or AI development experience.
- Strong Python programming skills.
- Experience with PyTorch and Hugging Face Transformers.
- Experience deploying open-source LLMs locally or on private infrastructure.
- Good understanding of LLM architecture and Generative AI concepts.
- Experience working in Linux environments.
- Experience with Docker and Git.
- Strong analytical and problem-solving skills.
- Passion for learning new AI technologies.
- Technical Skills (Preferred)
Experience with one or more of the following technologies is an advantage:
- LoRA / QLoRA
- Model Fine-Tuning
- Model Quantisation
- GGUF
- Ollama
- vLLM
- CUDA
- Kubernetes
- Vector Databases
- LangChain / LangGraph
- Model Context Protocol (MCP)
- AI Agents
- Retrieval-Augmented Generation (RAG)
- MLflow
- Model Evaluation
- Performance Benchmarking
What You'll Be Building
You will contribute to projects such as:
- Enterprise AI Platform
- Domain-specific AI Models
- AI Model Serving Infrastructure
- AI Evaluation Frameworks
- Intelligent AI Agents
- AI Knowledge Platforms
- AI-powered Developer Platform
- Enterprise AI APIs
Why Join Us?
- Build proprietary AI capabilities instead of only consuming AI services.
- Work with the latest open-source AI technologies.
- Influence the company's long-term AI strategy.
- Collaborate with an experienced engineering team.
- Continuous learning and technical growth opportunities.
- Flexible remote working environment.
- Opportunity to build AI products with real business impact.
Don’t miss this opportunity to accelerate your AI engineering career and become part of our growing AI team.
