About this role
Applied ML Engineer at Foxglove. Location: San Francisco, California, United States. Role: Deploy inference, Build embeddings, Design training Requirements: Production ML infrastructure, cloud inference, model serving optimization, vector databases, and collaboration with engineers. Category: Engineering Seniority: Senior Level Tools: TorchServe, vLLM, Triton, Pinecone, Lance, turbolpuffer, pgvector, AWS, GCP Commitment: Full Time Workplace: Onsite Languages: English