본문으로 건너뛰기
← 전체 뉴스
MLflow Blog2026년 1월 29일

MLflow에 DeepEval, RAGAS, Phoenix Judges 도입

영어 원문을 AI가 번역했습니다. 영어로 보기

요약

MLflow has integrated over 50 LLM-as-a-judge metrics from DeepEval, RAGAS, and Arize Phoenix directly into the MLflow Scorer API. Practitioners can now run and compare these third-party evaluations alongside native MLflow judges in a single API call and unified visual UI to debug and evaluate RAG pipelines and agents.

MLflow의 광범위하고 업계 최고의 고품질 LLM 심사위원 스위트를 사용하여 에이전트를 개선하세요.

관련 기사

News

Jev가 LLM judge를 대체할 수 있을까요? 파트 2: 더 어려운 답변 테스트

mlflow-blog7d ago
News

Jev가 LLM judge를 대체할 수 있을까? 품질, 비용, 지연 시간 평가

mlflow-blog10d ago
News

MLflow를 활용한 AI 에이전트 스킬 평가 및 개선

mlflow-blogAug 2026
News

리뷰 큐: 더 나은 AI를 향한 인간의 개입 단계

mlflow-blogJul 2026