-
GenAI Evaluation with mlflow.genai.evaluate(): Beyond Accuracy
How to evaluate RAG and agentic GenAI systems with MLflow 3.x, mlflow.genai.evaluate(), LLM-as-a-judge scorers, custom trace-aware metrics, and prompt A/B tests.
10 min read -
Productionising Generative AI with MLflow 3.x: Tracing, Evaluation, and Prompt Optimisation
How MLflow 3.x helps productionise GenAI systems with OpenTelemetry-compatible traces, trace-aware evaluation, RAG judges, custom scorers, and prompt registry workflows.
12 min read -
MkDocs RAG Documentation Assistant
A practical look at building a MkDocs documentation assistant with RAG, FastAPI, Gemini, vector search, cited answers, and a deployable documentation workflow.
9 min read