This role may have closed.
Listings older than 30 days are deprioritized in search. The link to the source is still below in case the posting is still open. For active openings, see similar jobs.
DeepX
LLM Serving SW Engineer
Seoul, South KoreaPosted 13 months ago
What you'd do
- Develop and optimize LLM serving systems based on DeepX NPU hardware.
- Design and implement runtime and inference engines for LLM models.
- Analyze LLM serving performance and resolve bottlenecks.
What they want
- Bachelor's degree or higher in Computer Science or related field.
- Software development experience using C/C++ and Python.
- Understanding of LLMs or deep learning frameworks such as TensorFlow or PyTorch.
Nice to have
- Experience with AI accelerator hardware such as NPUs or GPUs.
- Hands-on experience with model compilers/runtimes like ONNX,TVM,or TensorRT.
- Knowledge of model optimization techniques such as quantization and pruning.
Seoulstart's read
- Visa tier (likely)
- E-7 specialty / professional (likely)
- Korean language
- Korean not required
- English friendliness
- High: JD in English, foreign-friendly signals