Splunk Blogs

.conf26

Cisco Data Fabric

Artificial Intelligence

Latest Articles

A New Way to Work with Splunk: Splunk CLI and Agent Skills
Artificial Intelligence
7 Minute Read

A New Way to Work with Splunk: Splunk CLI and Agent Skills

We’re introducing splunkctl, the new Splunk CLI purpose built for the agentic era, and expanding the library of Splunk Agent Skills.
Stop Guessing What Matters: Move Beyond Golden Signals to Curated Business Journeys
Observability
7 Minute Read

Stop Guessing What Matters: Move Beyond Golden Signals to Curated Business Journeys

We’re introducing Business Journeys in Splunk Observability Cloud.
Choosing an Embedding Model for RAG: Why Your Custom Evals Beat the Public Leaderboard
Artificial Intelligence
10 Minute Read

Choosing an Embedding Model for RAG: Why Your Custom Evals Beat the Public Leaderboard

Public embedding leaderboards help shortlist models, but only corpus-specific evaluations confirm fit. Learn how to measure performance.
Optimizing RAG Retrieval: How to Select the Right Reranking Model
Artificial Intelligence
11 Minute Read

Optimizing RAG Retrieval: How to Select the Right Reranking Model

Rerankers optimize retrieval order for production RAG systems; learn how to select reranker architectures for your latency budget, and domain requirements.
Protecting Critical Apps in the Age of AI Agents
.conf
10 Minute Read

Protecting Critical Apps in the Age of AI Agents

How Splunk connects runtime and security context for faster response.
LLM-as-Judge vs. Human Evaluation: When to Use Each (And Why Elite Teams Use Both)
Learn
7 Minute Read

LLM-as-Judge vs. Human Evaluation: When to Use Each (And Why Elite Teams Use Both)

Understand the tradeoffs between LLMs and humans for generative AI evaluation
LLM Judges vs. SLM Judges: When To Use Which
Learn
8 Minutes Read

LLM Judges vs. SLM Judges: When To Use Which

LLM judge vs SLM judge: an SLM costs 10-30x less and runs in 15-150ms, making 100% coverage affordable. See where each wins and when to make the switch.
Benchmarks for Multi-Agent AI Systems
Learn
5 Minute Read

Benchmarks for Multi-Agent AI Systems

Evaluate multi-agent AI systems using benchmarks that prioritize coordination, reliability, and policy adherence over simple accuracy scores to ensure production readiness.
SLM-as-Judge: How to Build and Deploy an SLM Judge
Learn
7 Minute Read

SLM-as-Judge: How to Build and Deploy an SLM Judge

How to build an SLM judge: See how to size the model, fine-tune with LoRA, validate it, and serve it in production.