๐ŸŽฏ

deep-agents-orchestration

๐ŸŽฏSkill

from langchain-ai/skills-benchmarks

VibeIndex|
What it does
|

A benchmarking framework that measures how skill documentation design affects Claude Code's adherence to recommended patterns, with support for configurable task treatments and parallel test execution in Docker sandboxes.

๐Ÿ“ฆ

Same repository

langchain-ai/skills-benchmarks(21 items)

deep-agents-orchestration

Installation

Vibe Index InstallInstalls to .claude/skills/
npx vibeindex add langchain-ai/skills-benchmarks --skill deep-agents-orchestration
skills.sh Installโš  Installs to .agents/skills/
npx skills add langchain-ai/skills-benchmarks --skill deep-agents-orchestration
Manual InstallCopy SKILL.md content and save to the path below
~/.claude/skills/deep-agents-orchestration/SKILL.md

SKILL.md

46Installs
-
AddedMar 12, 2026

More from this repository10

๐ŸŽฏ
react-components๐ŸŽฏSkill

A benchmarking framework from LangChain that measures how skill documentation design affects Claude Code's adherence to recommended patterns, using Docker-sandboxed tasks with treatment-based test variations.

๐ŸŽฏ
langchain-oss-primer๐ŸŽฏSkill

A benchmark suite by LangChain that measures how skill documentation design affects Claude Code's adherence to recommended patterns, using Docker-sandboxed tasks with configurable treatments and repetitions.

๐ŸŽฏ
langchain-dependencies๐ŸŽฏSkill

A benchmarking framework that measures how skill documentation design affects Claude Code's adherence to recommended patterns. It supports multiple treatments, repetitions, and parallel test execution with Docker-sandboxed validation.

๐ŸŽฏ
langgraph-persistence๐ŸŽฏSkill

Part of LangChain's skill benchmarks project that measures how skill documentation design affects Claude Code's adherence to recommended patterns, using Docker-sandboxed test tasks with configurable treatments.

๐ŸŽฏ
framework-selection๐ŸŽฏSkill

A benchmarking framework that measures how skill documentation design affects Claude Code's adherence to recommended patterns, using Docker-sandboxed tasks with configurable treatments and automated validation.

๐ŸŽฏ
api-docs๐ŸŽฏSkill

A benchmark skill from LangChain that measures how skill documentation design affects Claude Code's adherence to recommended patterns, with support for multiple treatments and configurable model selection.

๐ŸŽฏ
database-migrations๐ŸŽฏSkill

A benchmarking framework that measures how skill documentation design affects Claude Code's adherence to recommended patterns, using Docker-sandboxed tasks with configurable treatments and validation scripts.

๐ŸŽฏ
langsmith-trace๐ŸŽฏSkill

A benchmarking framework that measures how skill documentation design affects Claude Code's adherence to recommended patterns. Supports multiple treatments, repetitions, and parallel test execution in Docker-sandboxed environments.

๐ŸŽฏ
deep-agents-core๐ŸŽฏSkill

A benchmarking framework that measures how skill documentation design affects Claude Code's adherence to recommended patterns. It runs tasks in Docker-sandboxed environments with configurable treatments and repetitions to evaluate skill effectiveness.

๐ŸŽฏ
testing-patterns๐ŸŽฏSkill

A benchmarking framework from LangChain that measures how skill documentation design affects Claude Code adherence to recommended patterns, using configurable treatments, Docker-sandboxed execution, and automated validation.