New in Arize AX: first-class sessions, Agent-as-a-Judge, and vision evals Arize AI shipped sessions as a first-class unit of work in Arize AX between August 6 and September 18, letting users annotate, queue, preview, and filter whole conversations rather than individual spans. The release also makes Agent-as-a-Judge available on every plan, adds Claude Fable 5.1, GPT-6 Astra, Gemini 3.7 Flash, and Mistral models on AWS Bedrock to Playground and evaluators, and enables LLM-as-a-judge templates to reference image columns for vision evals. Dataset population from traces now writes every matching record instead of applying a sampling budget and appends 2-7x faster on every account tier. Here’s everything that landed in Arize AX https://arize.com/products/ax/ between August 6 and September 18. The main point: sessions are now a first-class unit of work. You can annotate a whole conversation, queue it for review, preview it in an evaluator https://arize.com/glossary/evaluations/ , and ask Alyx https://arize.com/products/alyx/ to filter for it in plain language. Everything else is below, and you can explore our full Changelog to see everything we’re shipping https://arize.com/docs/ax/release-notes . Oh, and since it’s Jev week: you can read about our latest thoughts around if decision models can replace LLM judges on our blog https://arize.com/blog/typesafe-jev-llm-judge/?utm source=hs email&utm medium=email& hsenc=p2ANqtz-86CIZRCWMERIs6O74KYIMCuhinRObsaEPaoaHGpU 2d1eNU5N1 w0ANCiQ3ppjMPXpOw4C and expect more soon . Let’s jump in. Work with whole sessions, not just spans Sessions now behave like a single object you can act on. Annotate one from its detail view and the label is stored as session annotation.