Skip to content
MLflowChangelog · September 2026

MLflow 3.16 made the redesigned trace explorer default and patched basic-auth security.

MLflow released version 3.16.0, making the redesigned trace explorer its default interface [16]. The release added session grouping for multi-turn conversations, span links, and per-user budget policies in the AI Gateway [16]. MLflow 3.16.1 followed shortly after to patch a hardcoded administrative password in basic authentication, while also adding span link support for Unity Catalog traces and resolving trace location lookups in Model Serving [11]. In parallel, MLflow 2.11.5 introduced an option to route Unity Catalog model registry artifact transfers through the Databricks SDK Files API [7].

LLM evaluation and agent tracking remained central throughout the month. Production case studies demonstrated evaluation-first architectures that paired traces with automated judges to maintain quality gates [13], while practitioners documented patterns for tracking per-run agent costs with LangGraph [10]. Other evaluations tested whether specialized evaluation models like Jev could cut latency and compute spend for MLflow question-answering assessments relative to general-purpose frontier models [4] [8].

Community members faced operational hurdles around experiment management and trace persistence. Practitioners worked through behavioral differences between Workspace and Unity Catalog experiment targets when using autologging [17]. At the end of the month, users reported an issue where StartTraceV3 returned success codes but left traces unretrievable via GetTrace [1].

Everything cited

  1. [1]MLflow traces accepted by StartTraceV3 (200 OK) but never stored: GetTrace returns NOT_FOUND community · 2026-09-30
  2. [2]v1.19.0 release · 2026-09-30
  3. [3]Workshop on bringing systematic evaluation and MLflow tracking to LLM apps, Oct 3 community · 2026-09-28
  4. [4]Can Jev replace your LLM judge? Part 2: Testing harder answers news · 2026-09-25
  5. [5]Laya off the benchmark: can a zero-shot decision model route real SQL traffic? community · 2026-09-24
  6. [6]v1.18.0 release · 2026-09-24
  7. [7]MLflow 2.11.5 release · 2026-09-23
  8. [8]Can Jev replace your LLM judge? Evaluating quality, cost, and latency news · 2026-09-22
  9. [9]How adidas Uses Databricks to Build Better Products video · 2026-09-21
  10. [10]Building Custom Agents on Databricks:LangGraph, Atlan-Grounded Routing, and Per-Run Cost with MLflow community · 2026-09-17
  11. [11]MLflow 3.16.1 release · 2026-09-17
  12. [12]Could a “Data → Agent” composer be useful for Databricks? community · 2026-09-15
  13. [13]Evaluation-First AI Agents: How Zepto Scales Customer Support on Databricks and MLflow news · 2026-09-09
  14. [14]How to Build and Serve Production ML Features | Databricks Feature Store Demo video · 2026-09-08
  15. [15]Building Custom Agents on Databricks:LangGraph, Atlan-Grounded Routing, and Per-Run Cost with MLflow community · 2026-09-05
  16. [16]MLflow 3.16.0 release · 2026-09-04
  17. [17]Difference between Workspace and Unity Catalog experiments when using MLflow autologging? community · 2026-09-02

A frozen monthly snapshot, generated from the brickster.ai archive and never rewritten. For the live view of this topic, see the MLflow hub. brickster.ai is an independent community project, not affiliated with Databricks, Inc.