Test pyramid
HoldTechniques
A layered testing strategy that balances unit, integration, and higher-level tests.
Why it's here
Placed in Hold: 1 article(s) of evidence from 1 source(s), led by research-stage coverage, with 0 in the last 30 days. Confidence 24%. Low accumulated evidence, so it defaults conservatively pending more signal.
Evidence (1)
- 4InfoQ·7/7/2026researchDesigning Reliable AI Platforms
Aaron Erickson describes how NVIDIA designs and tests AI agent hierarchies for production use. The presentation emphasizes combining deterministic tools with agentic discovery, using rare context effectively, and applying LLM-as-a-judge test pyramids to improve reliability at scale.