Source

langwatch/skills

23 skills · 932 combined installs

Skills from this source

#
Skill
Source
8W Activity
Installs
1
scenarios Test your AI agent with simulation-based scenarios. Covers writing scenario test code (Scenario SDK), creating platfo…
langwatch/skills
150
2
tracing Add LangWatch tracing and observability to your code. Use for both onboarding (instrument an entire codebase) and tar…
langwatch/skills
137
3
evaluations Compatibility router for LangWatch evaluation requests.
langwatch/skills
120
4
level-up Take your AI agent to the next level with full LangWatch integration. Adds tracing, prompt versioning, evaluation exp…
langwatch/skills
111
5
prompts Version and manage your agent's prompts with LangWatch Prompts CLI. Use for both onboarding (set up prompt versioning…
langwatch/skills
103
6
datasets Generate realistic synthetic evaluation datasets by analyzing the user's codebase, prompts, production traces, and re…
langwatch/skills
68
7
analytics Analyze your AI agent's performance using LangWatch analytics. Use when the user wants to understand costs, latency, …
langwatch/skills
67
8
online-evaluations Configure LangWatch online evaluations and guardrails for production traffic. Use when the user wants to score live t…
langwatch/skills
27
9
experiments Create and run LangWatch experiments for pre-deployment batch testing. Use when the user wants to test an agent again…
langwatch/skills
25
10
agent-performance Deep-dive diagnosis of how your AI agent behaves in production. Explores LangWatch analytics and traces end to end to…
langwatch/skills
20
11
debug-instrumentation Debug and improve your LangWatch traces. Inspects production traces for missing input/output, disconnected spans, unl…
langwatch/skills
16
12
evaluate-multimodal Evaluate multimodal AI agents that process images, audio, PDFs, or other files. Sets up evaluations using LangWatch's…
langwatch/skills
16
13
generate-rag-dataset Generate a synthetic evaluation dataset from your RAG knowledge base. Creates diverse Q&A pairs with expected answers…
langwatch/skills
14
14
improve-setup Expert AI engineering consultant for your LangWatch setup. Audits your codebase, traces, evaluations, and scenarios, …
langwatch/skills
12
15
test-cli-usability Write scenario tests that verify your CLI tool is usable by AI agents. Ensures commands work non-interactively, provi…
langwatch/skills
11
16
test-compliance Test that your AI agent stays observational and doesn't give prescriptive advice in regulated domains (healthcare, fi…
langwatch/skills
11
17
debug-with-langwatch Root-cause production errors and misbehaving agent runs with LangWatch. Finds errored traces, inspects spans, checks …
langwatch/skills
6
18
agent-best-practices Expert AI engineering consultant for your agent development practices. Audits your codebase, traces, evaluations, and…
langwatch/skills
4
19
connect-agent Connect the codebase's AI agent to LangWatch agent simulations, so test suites run against the real agent process. Ad…
langwatch/skills
4
20
eval-triage Investigate failing experiments and evaluations with LangWatch. Triage a failing experiment run to the exact rows and…
langwatch/skills
4
21
setup-lw Set up and troubleshoot the LangWatch CLI, covering login (cloud and self-hosted), endpoint configuration, project se…
langwatch/skills
4
22
context-sweet-spot Investigates the context economics of your own coding-agent sessions in LangWatch. Reads real sessions to find where …
langwatch/skills
1
23
provider-cost-comparison Prices your real LangWatch usage mix against other model providers. Exports your actual token mix per model, includin…
langwatch/skills
1