Nemo Relay Tune Performance
NVIDIA/NeMo-Relay/skills/nemo-relay-tune-performanceai-mlOfficial
Official Provider SkillView repo
Plan a measured NeMo Relay adaptive tuning rollout after baseline scopes, tool calls, LLM calls, and observability are working; use this skill to improve latency, tool parallelism, prompt-cache behavior, or model-request behavior from runtime signals
Files2 files
SKILL.md1 lines
Loading editor…
Install
RecommendedOne command — your agent picks it up automatically.
Select an AI agent above to see the install command.
or
Manual Install
More stepsDownload the archive and add the files to your project manually.
Skill details
Versionv1.0.0
AuthorNVIDIA
Categoryai-ml
Skill IDNVIDIA/NeMo-Relay/skills/nemo-relay-tune-performance
Files2 files
Related skills
Add Binding FeatureAdd or change a public NeMo Relay API surface across the core runtime and every affected bindingAdd MiddlewareAdd a new guardrail or intercept type to the NeMo Relay middleware pipelineContribute ApiContribute a new NeMo Relay public API surface safely, with binding parity and docs in mindContribute DocsContribute documentation or example changes that stay aligned with NeMo Relay public behaviorContribute IntegrationContribute a new or updated third-party framework integration for NeMo RelayKarpathy GuidelinesBehavioral guidelines to reduce common LLM coding mistakes. Use when writing, reviewing, or refactoring code to avoid overcomplication, make surgical changes, surface assumptions, and define verifiable success criteria.