Jobs Companies Nanyang Technological University AI Engineer, Platforms

About this AI Engineer, Platforms role at Nanyang Technological University

Nanyang Technological University · Onsite · NTU Main Campus, Singapore

AI Singapore (AISG) is a national AI programme launched by the National Research Foundation (NRF), Singapore, to build and anchor deep national capabilities in AI. AISG is supported through a government-wide partnership including the NRF, Ministry of Digital Development and Information (MDDI), Infocomm Media Development Authority (IMDA), Economic Development Board (EDB) and Enterprise Singapore (ESG). We bring together research institutions and the vibrant ecosystem of AI start-ups and companies to support impactful research, develop talent, and power Singapore's AI efforts.

This position will be hosted at the Nanyang Technological University (NTU) under VP (Artificial Intelligence & Digital Economy)’s office and we welcome you to join our community.

We're looking for an AI Engineer to join the Platform team within AI Products at AISG. In this role, you will be collaborating with different internal teams to design and implement optimized inference workflows and support the team to build customized LLM-based solutions.

Your work will directly contribute to the deployment, optimisation and management of large language models (LLMs) in production, retrieval-augmented generation (RAG) services, AI agent orchestration platforms and GPU-enabled AI infrastructure.

Responsibilities:

Platform operations and reliability

  • Own day-to-day operations of SEA-LION API Farm, our multi-cloud LLM inference platform — monitoring environment health, GPU capacity, performance, cost, and security posture.

  • Optimise LLM inference across various modalities to drive business value and support production goals.

  • Diagnose and troubleshoot performance and reliability issues on API Farm.

  • Build new API services such as batch API services, MCP services.

Infrastructure, CI/CD, and automation

  • Manage high performance AI clusters and storage systems using infrastructure-as-code (e.g. Terraform) across different cloud providers.

  • Develop and maintain CI/CD pipelines, container build/registry workflows, and deployment automation so teams can ship safely and frequently.

  • Strengthen observability across the stack including logs, metrics, traces, and dashboards, and reduce toil by automating repetitive operational tasks.

AI-assisted ops and continuous improvement

  • Use AI tools (e.g. Claude, Copilot, Cursor) appropriately in your daily work responsibilities.

  • Build internal tools leveraging AI to reduce manual effort in day-to-day operations.

Requirements:

You should be a hands-on engineer who is comfortable operating cloud and GPU infrastructure end-to-end, who understands how to deploy and run large language models reliably in production, and who actively uses AI tools to make platform work faster and more reliable.

  • A degree in Computer Science, Information Technology, or equivalent.

  • 1–3 years of DevOps, SRE, or platform engineering experience, with a track record of operating production systems at scale.

  • Hands-on experience operating workloads on different cloud providers including IaC (e.g. Terraform), containers and orchestration (e.g. Docker, Kubernetes), and managed services for compute, storage, and networking.

  • Strong knowledge on Inference frameworks and libraries (e.g., vLLM, SGLang, TensorRT-LLM, Transformers).

  • Hands-on experience deploying and serving LLMs in production — model serving, GPU scheduling, autoscaling, latency/throughput optimisation, and inference cost management.

  • REST API design, model context protocol (MCP), Internet authentication patterns (e.g. OAuth).

  • Strong fundamentals in CI/CD, observability (logs/metrics/traces), and incident response.

  • Demonstrated use of AI tools (e.g. Claude, Copilot, Cursor) in your day-to-day engineering — for code generation, review, debugging, and documentation — with a clear sense of where they help and where they don't.

  • Solid scripting/programming skills (e.g. Python, Bash) and comfortable reading other people's code across the stack.

  • Strong communication skills with the ability to explain technical concepts.

Good to Have:

  • Experience with multimodal AI models (e.g. vision language models, audio language models).

  • C/C++/Rust/Go or other relevant programming languages.

  • Contributions to open-source AI/ML projects.

We regret that only shortlisted candidates will be notified.

Hiring Institution: NTU

Ready to apply to Nanyang Technological University?
Apply to Nanyang Technological University

About Nanyang Technological University

A research-intensive public university, Nanyang Technological University, Singapore (NTU Singapore) has 33,000 undergraduate and postgraduate students in the Engineering, Business, Science, Humanities, Arts, & Social Sciences, and Graduate colleges. It also has a medical school, the Lee Kong Chian School of Medicine, established jointly with Imperial College London. NTU is also home to world-class autonomous institutes – the National Institute of Education, S Rajaratnam School of International Studies, Earth Observatory of Singapore, and Singapore Centre for Environmental Life Sciences Engineering – and various leading research centres such as the Nanyang Environment & Water Research Institu

See all jobs at Nanyang Technological University →

Similar jobs

Nanyang Technological University
AI Engineer, Dev Ops
Nanyang Technological University
⚡ Apply early NTU Main Campus, Singapore Onsite
● New 👁 Seen ✓ Applied 1w ago
Nanyang Technological University
Research Engineer I (Computer Science / Artificial Intelligence / Data Science / Electrical Engineering)
Nanyang Technological University
⚡ Apply early NTU Main Campus, Singapore Onsite
● New 👁 Seen ✓ Applied 1w ago
Nanyang Technological University
AI Engineer, Full Stack
Nanyang Technological University
⚡ Apply early NTU Main Campus, Singapore Onsite
● New 👁 Seen ✓ Applied 1w ago
TwelveLabs
Staff ML Research Engineer, Multimodal Structure & Marengo
TwelveLabs
⚡ Apply early Seoul, South Korea Hybrid
● New 👁 Seen ✓ Applied 5h ago
Ruby Labs
Senior AI Engineer
Ruby Labs
⚡ Apply early European Union · location restricted
● New 👁 Seen ✓ Applied 5h ago
Adelphi
Machine Learning Engineer
Adelphi
⚡ Apply early Washington D.C. · location restricted
● New 👁 Seen ✓ Applied 5h ago
Adelphi
Senior ML Engineer
Adelphi
⚡ Apply early Remote · location restricted
● New 👁 Seen ✓ Applied 5h ago
Winamax
Senior ML / AI Engineer - Team Data (H/F)
Winamax
⚡ Apply early Paris Onsite
● New 👁 Seen ✓ Applied 5h ago
Spotify
Senior Staff Machine Learning Engineer - Content Platform
Spotify
⚡ Apply early New York, NY · location restricted
● New 👁 Seen ✓ Applied 5h ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at Nanyang Technological University

See all jobs at Nanyang Technological University →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free