About this role
LLM Engineer (LLM Evaluation) at 42dot. Location: Pangyo or Korea. Role: Design benchmark, Automate evaluation, Validate models Requirements: Experience in LLM evaluation and benchmarking, design of evaluation protocols, automation pipelines, and end-to-end workflow integration using Kubernetes, Argo, MLflow; strong Python service development and focus on reproducibility. Category: Research and Development (R&D) Seniority: Mid Level Tools: Kubernetes, Argo Workflows, MLflow, Python, Async programming Commitment: Full Time Workplace: Hybrid Languages: English