aijobs.net

Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start

Singapore, Singapore

SGD 66K-78K (estimate) Entry-level Full Time

Apply Save
Found 5h ago
Tasks
Perks/Benefits
Skills/Tech-stack

Activation functions | C# | C++ | CUDA | Computation Graph | Computation graph optimization | Constant folding | Deep learning | Distributed inference | GPU Memory Model | GPU Performance | GPU Programming | GPU memory | GPU performance analysis | Graph optimization | Matrix Operations | Memory Reuse | Memory model | Mixture of Experts | Normalization | Nsight | Operator fusion | Performance Analysis | Pipeline parallelism | Precision alignment | Profiling | Python | Quantization | SGLang | Scheduling optimization | Sequence parallelism | Tensor Parallelism | TensorRT-LLM | VLLM | Vectorization

Education

Bachelor of Engineering | Bachelor of Science | Master of Science

Roles

Backend | Backend Inference Runtime Engineer | Engineer | Runtime Engineer

Regions

Asia/Pacific

Countries

Singapore

Cities

Singapore, SG

Apply Save
Language: en Views: 1 Clicks: 1 Saves: 0

Related jobs