Jackson Wells' Blog Posts

Jackson Wells

Jackson Wells is a writer and marketer at Splunk focused on how teams evaluate and monitor AI agents in production. He came to Splunk through the Galileo acquisition, and before that spent four years writing for developer audiences at LinearB. His work covers evals, guardrails, and what it takes to trust a AI systems in production.

How to Evaluate AI Systems
Learn
10 Minute Read

How to Evaluate AI Systems

Explore a detailed step-by-step process on effectively evaluating AI systems to boost their potential.
LLM Benchmarks: Top Categories for Evaluating AI Beyond Conventional Metrics
Learn
6 MINUTE READ

LLM Benchmarks: Top Categories for Evaluating AI Beyond Conventional Metrics

Evaluating LLMs requires moving beyond general metrics to domain-specific, agentic, and adversarial benchmarks that reflect real-world performance.
How To Benchmark Autonomous AI Agents Effectively
Artificial Intelligence
7 minute read

How To Benchmark Autonomous AI Agents Effectively

Evaluate autonomous agents through a framework that prioritizes multi-step decision paths, tool interaction accuracy, and production-ready safety guardrails.