This role may have closed.
Listings older than 30 days are deprioritized in search. The link to the source is still below in case the posting is still open. For active openings, see similar jobs.
Upstage
AI Research Engineer - LLM Eval
Seoul, South KoreaPosted 2 months ago
What you'd do
- Research and develop LLM evaluation benchmarks for knowledge,reasoning,alignment,and agentic tool use
- Monitor global LLM benchmark trends and new evaluation methodologies
- Develop multilingual evaluation datasets
What they want
- Strong background in NLP or ML research
- Experience with LLM evaluation or benchmark development
- Programming proficiency for building evaluation infrastructure
Nice to have
- International AI publications strongly preferred
- Experience with multilingual NLP
- Familiarity with distributed computing for ML experiments
Seoulstart's read
- Visa tier (likely)
- E-7 specialty / professional (likely)
- Korean language
- Business-level Korean required
- English friendliness
- Medium: English JD, mixed-language workplace
- Context
- Upstage AI Research Engineer for LLM Eval of Solar LLM; fully remote (Anywhere on Earth option available); international publications preferred.