Find jobs in AI/ML, Data Science and Big Data
11 results
for FlashAttention
(Skill/Tech stack)
-
Model Optimization Engineer USD 100K-150KC++ | CUDA | Continuous batching | Deep learning | DeepSpeedSenior-level Full TimeUnited States - Remote R1d ago
-
ML Performance Engineer USD 100K-150KBenchmarking | C++ | Continuous batching | Cutlass | Deep learningCareer growth | Direct W2 employment | Remote workSenior-level Full TimeTempe, AZ R1d ago
-
Entry-level Full Time北京 R4d ago
-
AI Optimization Engineer USD 100K-150KBenchmarking | C++ | Cache optimization | Compiler optimization | Continuous batchingCareer growthSenior-level Full TimeUnited States - Remote R4d ago
-
Senior-level Full TimeUnited States - Remote R5d ago
-
Senior-level Full TimeChina, Shanghai6d ago
-
Mid-level Full Time北京 R6d ago
-
AI Engineer (Managed Services) SGD 85K-138KARES | AWQ | Agent Orchestration | Agent systems | Attention MechanismsMid-level Full TimeSingapore11d ago
-
Mid-level Full TimeSingapore13d ago
-
Lead Machine Learning Engineer, Inference & Performance USD 184K-270KAttention Mechanism | Batching | CUDA | Data Engineering | FlashAttention401k employer match | Bereavement leave | Comprehensive health insurance | Employee referral bonuses | Paid HolidaysSenior-level Full TimeRemote R19d ago
-
Bayesian optimization | Data Generation | Debugging | DeepSpeed | Distributed SystemsAdditional time off for learning and development | Annual leave | Cycle to work scheme | Employee assistance program | Group personal pensionEntry-level ContractLondon, United Kingdom1mo ago