# @techbible.ai on Instagram

- **Type:** Video
- **Original URL:** https://www.instagram.com/p/DalCFRKogJC
- **Gondola URL:** https://gondola.cc/posts/67549451-techbibleai-instagram
- **Thumbnail:** https://img.gondola.cc/tr:w-,h-,fo-auto/postThumbnails/8f7087f82f.jpg
- **Posted:** 2026-07-09T16:24:48.000+00:00
- **Account Owner:** TechBible | Ghita El Haitmy (@techbible.ai) — https://gondola.cc/techbible.ai

## Caption

Testing isn’t a one-time event 👇 

 (Logging): You cannot test what you don’t see. You must log every response your agent generates in production. Tools like braintrustdev help you capture these traces persistently.

Evaluate (Scoring): Run your logs against your chosen evaluation method (e.g., have an LLM grade the logs against a rubric) to find exactly where the model fails.

Iterate (Improvement): Once you find the failure, you fix the prompt, the skill, or the harness configuration, and then re-run the evaluation to confirm the fix works without breaking something else.

Follow for more AI tips and updates 🚀

#aiagents #evals

## Stats

- **Views:** 0
- **Likes:** 137
- **Shares:** 0
- **Comments:** 24

## Tags

aiagents, evals

---
Copyright (c) Gondola