aijobs.net

Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start

Singapore, Singapore

SGD 60K-65K (estimate) Entry-level Full Time

Apply Save
Found 3h ago
Tasks
Perks/Benefits
Skills/Tech-stack

Asynchronous Scheduling | Benchmarking | Cache optimization | Compilation Optimization | Distributed Systems | GPU Computing | GPU Memory Optimization | GPU Performance | GPU Performance Optimization | GPU memory | High Performance | High-Performance Computing | Language Models | Large Language Models | Machine Learning | Memory Optimization | Mixture of Experts | NPU | Operator fusion | Performance Computing | Performance optimization | Pipeline parallelism | Sequence parallelism | Tensor Parallelism | TensorRT-LLM | VLLM

Education

Bachelor of Engineering | Bachelor of Science | Master of Science

Roles

Backend | Backend Inference Runtime Engineer | Engineer | Runtime Engineer

Regions

Asia/Pacific

Countries

Singapore

Cities

Singapore, SG

Apply Save
Language: en Views: 1 Clicks: 0 Saves: 0

Related jobs