The role
Job description
Build LLM as judge rubrics; Correlate audio stt llm tool and tts traces; Create adversarial conversational datasets; Design audio native evaluation frameworks; Generate fine tuning data and prompt updates; Implement end-to-end observability; Mine production traces for failure patterns; Validate fixes using adversarial replay;
Index terms