Senior AI Agent Engineer
hace 14 días
Stand up eval suites using various evaluation frameworks and tooling, included but not limited to. Promptfoo, Braintrust, LangSmith, DeepEval, LLM-as-judge methods, and custom harnesses. Build the citations and confidence.