About this role
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco at Plaud Inc.. Location: San Francisco, California, United States. Role: building dashboards, defining benchmarks, monitoring health Requirements: Python software engineering; building distributed systems, data pipelines, and evaluation harnesses at scale; partner with ML researchers to define benchmarks; build dashboards and monitor model health; debug mid-training anomalies; communicate results clearly. Category: Engineering Seniority: Mid Level Tools: Python, Distributed systems, Data pipelines, Evaluation harnesses, Dashboards, Weighs & Biases, MLflow Commitment: Full Time Workplace: Hybrid Languages: English