This role may have closed.
Listings older than 30 days are deprioritized in search. The link to the source is still below in case the posting is still open. For active openings, see similar jobs.
Nota AI
AI Inference Optimization Engineer
Seoul, South KoreaPosted 2 months ago
What you'd do
- Analyze mismatches between AI models and target hardware (NPU,GPU,CPU) and track root causes to resolution
- Design and implement graph-level optimization passes (op fusion,folding,decomposition,replacement)
- Explore and experiment with optimization opportunities considering both model structure and hardware characteristics
What they want
- 2+ years relevant post-graduation practical experience,or Masters degree or higher
- PyTorch,ONNX,Python,Linux,Git practical experience
- End-to-end experience following AI model training through actual device execution
Nice to have
- Graph IR-level model analysis and transformation (TensorRT,TFLite,ExecuTorch)
- NPU or custom accelerator target model optimization experience
- Deep understanding of GPU/NPU kernels
Seoulstart's read
- Visa tier (likely)
- E-7 specialty / professional (likely)
- Korean language
- Business-level Korean required
- English friendliness
- Medium: English JD, mixed-language workplace
Related Seoulstart guides
Similar roles on Seoulstart
Solutions ArchitectDatabricks · Seoul, South Korea · 1 day ago
AI Engineer - FDE (Forward Deployed Engineer)Databricks · Seoul, South Korea · 1 day ago
Manager, Software Engineering, Publishing Engagement - Publishing PlatformRiot Games · Seoul, Korea · 1 day ago
Mobile Engineer - AI Finance AgentBjak · Seoul, Seoul, South Korea · 1 day ago