aijobs.net

Research Engineer - LLM/VLM Inference Optimization (Seed Infra)

Seattle, Washington, United States

USD 241K-456K Mid-level Full Time

Apply Save
Found 5h ago
Tasks
Perks/Benefits
Skills/Tech-stack

C# | C++ | CPU architecture | CUDA | Containerization | Conv2D | Cutlass | FlashAttention | GEMM | GEMV | GPU Architecture | GPU Programming | Graph Fusion | Low Precision | Low-precision computation | OpenCL | Parallel Computing | Performance Modeling | Performance Profiling | PyTorch | Python | Speculative decoding | Streaming inference | TensorFlow | TensorRT | Triton

Education

Bachelor of Engineering | Bachelor of Science | Master of Science

Roles

Engineer | Research Engineer

Regions

North America

Countries

United States

States

Washington, US

Cities

Seattle, Washington, US

Apply Save
Language: en Views: 1 Clicks: 0 Saves: 0

Related jobs