Splunk Blogs

.conf26

Cisco Data Fabric

Artificial Intelligence

Latest Articles

LLM-as-Judge vs. Human Evaluation: When to Use Each (And Why Elite Teams Use Both)
Learn
7 Minute Read

LLM-as-Judge vs. Human Evaluation: When to Use Each (And Why Elite Teams Use Both)

Understand the tradeoffs between LLMs and humans for generative AI evaluation
LLM Judges vs. SLM Judges: When To Use Which
Learn
8 Minutes Read

LLM Judges vs. SLM Judges: When To Use Which

LLM judge vs SLM judge: an SLM costs 10-30x less and runs in 15-150ms, making 100% coverage affordable. See where each wins and when to make the switch.
Benchmarks for Multi-Agent AI Systems
Learn
5 Minute Read

Benchmarks for Multi-Agent AI Systems

Evaluate multi-agent AI systems using benchmarks that prioritize coordination, reliability, and policy adherence over simple accuracy scores to ensure production readiness.
SLM-as-Judge: How to Build and Deploy an SLM Judge
Learn
7 Minute Read

SLM-as-Judge: How to Build and Deploy an SLM Judge

How to build an SLM judge: See how to size the model, fine-tune with LoRA, validate it, and serve it in production.
From Agentic Ops To Autonomous IT: The Next Leap in Technology Operations
Partners
6 Minute Read

From Agentic Ops To Autonomous IT: The Next Leap in Technology Operations

The conversation is shifting from AIOps to Agentic Ops.
How to Architect an Enterprise Retrieval-Augmented Generation (RAG) System
Artificial Intelligence
7 Minute Read

How to Architect an Enterprise Retrieval-Augmented Generation (RAG) System

Learn how to implement an end-to-end enterprise RAG architecture, step by step through the pipeline, by understanding common failures.
Born Observable: Why AI-Generated Apps and Agents Need Visibility from Day One
Observability
7 Minute Read

Born Observable: Why AI-Generated Apps and Agents Need Visibility from Day One

Observability Studio lets developers test and validate tracking data in real time during development, catching problems before they reach production.
How AI Coding Assistants Transform Security Teams into Tool Builders
Ciso Circle
5 Minute Read

How AI Coding Assistants Transform Security Teams into Tool Builders

Discover how security leaders use AI coding assistants and tiered sandboxes to empower frontline analysts to prototype faster without production risk.
Can AI Agents Catch Their Own Mistakes? Evaluating for Self-Correction
Artificial Intelligence
12 minutes

Can AI Agents Catch Their Own Mistakes? Evaluating for Self-Correction

Evaluate the effectiveness of AI agent self-reflection to determine when internal revision passes actually improve accuracy versus when they introduce new errors.