Job Title: Associate AI Platform Engineer
Bandar Seri Begawan , Brunei-M, Brunei Darussalam
Notice
- Closing Date: Tuesday, 06 October 2026
- Please submit a PDF copy of your CV, IC, Academic and Professional Certificate(s)
- Only shortlisted candidates will be notified
Job Description
The Associate AI Platform Engineer supports the operation, configuration, maintenance and continuous improvement of the organisation’s AI platform and related AI-runtime, accelerator, model-serving and application-platform services. The role performs routine platform administration, monitoring, troubleshooting, environment configuration, workload validation, deployment support and technical documentation in accordance with approved architecture, security controls and operating procedures. The role also develops practical capability in AI platform architecture, backend services, APIs, model serving, accelerator computing, containerisation, automation and platform engineering to progressively contribute to the enhancement and development of internal AI platform capabilities.
Responsibilities and Tasks
-
AI Platform Operations & Service Support:
- Manages the day-to-day operation and administration of assigned AI platforms
- Monitors platform performance, availability, usage, logs and alerts
- Sets up, configures, maintains and manages platform services and user environments
- Performs routine platform checks, updates and system validation
- Supports user access, configuration and technical service requests
- Identify platform issues and perform troubleshooting or escalate when required
- Coordinates with relevant technical teams to resolve issues involving systems, storage, networks, data or applications -
Monitoring, Incident Investigation & Troubleshooting:
- Performs initial troubleshooting for platform, API, container, runtime and model-related issues
- Reproduces reported issues and gathers relevant logs, errors, system details and user information
- Identifies which technical area is likely causing the issue
- Resolves routine issues using approved procedures and troubleshooting tools
- Escalates complex/high-risk issues to the appropriate technical team
- Tracks issues until resolved and document findings, actions and solutions
- Tests and confirms platform services are working properly after fixes or changes -
AI Runtime, Accelerator & Workload Support:
- Sets up and validates approved AI frameworks, libraries and system environments
- Supports the testing and running of AI training, inference and model workloads
- Monitors accelerator availability, usage, memory and workload performance
- Identifies basic compatibility, performance and resources issues
- Supports model-serving, inference services and API testing
- Works with AI/ML Engineers to ensure models and applications run effectively on the platform
- Documents workload requirements, test results, configuration issues and limitations -
Platform Configuration, Backend & Integration Support:
- Configures platform services, user access and application components
- Tests API, backend services, database connections and model-serving functions
- Supports configurations, scripting, automation and system integration changes
- Troubleshoots routine backend, API, authentication and connectivity issues
- Supports the integration of AI applications, models and platform services
- Understands how platform components, APIs, databases and user-facing systems work together
- Tests and validates new platform feature and service improvements -
Containerisation, Deployment & Automation Support:
- Builds, runs and tests containers and platform services using approved standards
- Manage container images, dependencies and deployment settings
- Supports routine deployments, restarts, rollbacks and service testing
- Deploys workloads using approved container and orchestration tools
- Develops or updates simple scripts for monitoring, testing and routine tasks
- Maintains configuration files,deployment records and technical assets
- Escalates complex deployment/platform issues to the appropriate technical team
Competencies
- AI Platform Architecture & Engineering
- AI Runtime Environment Setup
- Linux Server Administration
- AI Accelerator Computing
- Backend Services, API Integration & Model Serving
- Containerisation & Workload Orchestration
- Python Programming & Scripting
Area and Years of Experience
- Software development, AI/ML, platform engineering, cloud computing, systems administration or a related technical area with 1 year of experience
- Python, Linux, APIs, backend development, containers or AI-model development with no working experience required (working experience would be an advantage)
- Technical troubleshooting, application support or platform operations with no working experience required (working experience would be an advantage)
- AI platform, accelerator computing, model serving or cloud-based AI projects with no working experience required (working experience would be an advantage)
Education
Bachelor's Degree in Computer Science, Software Engineering, Information Technology, Artificial Intelligence, Data Science, Computer Engineering, Cloud Computing or relevant areas and equivalent working experience.