所有 Skills

找到 8642 个 Skills

Skills 列表

nemo-mbridge-perf-moe-comm-overlap

nemo-mbridge-perf-moe-comm-overlap

3Ktesting-qa

MoE expert-parallel communication overlap in Megatron Bridge. Covers dispatch/combine overlap, flex dispatcher backends, and expert wgrad scheduling.

nvidia avatarnvidia
获取
nemo-mbridge-resiliency

nemo-mbridge-resiliency

3Kbackend-api

Resiliency features in Megatron Bridge including fault tolerance, straggler detection, in-process restart, preemption, and re-run state machine.

nvidia avatarnvidia
获取
vss-setup-video-analytics-api

vss-setup-video-analytics-api

3Kdevops-cloud

Use to deploy the vss-video-analytics-api REST service standalone (config-source, data-log bind, Elasticsearch, optional Kafka). Not for full warehouse deploy.

nvidia avatarnvidia
获取
vss-manage-video-io-storage

vss-manage-video-io-storage

3Kbackend-api

Use to call the VIOS REST API (sensor list, timelines, clip extraction, snapshots, add/delete sensors and streams). Not for VLM inference or search.

nvidia avatarnvidia
获取
vss-deploy-video-embedding

vss-deploy-video-embedding

3Kdevops-cloud

Use this skill when deploying, operating, or integrating the VSS 3.2 GA RT-Embed Video Embedding microservice. Covers Docker Compose bring-up, GPU and storage prerequisites, the `/v1` REST API (file uploads, text and video embeddings, live RTSP streams, health and metrics), Redis/Kafka/OTel integration, common failure modes, and teardown.

nvidia avatarnvidia
获取
vss-deploy-detection-tracking-2d

vss-deploy-detection-tracking-2d

3Kbackend-api

Use this skill when the user wants to deploy, run, debug, tear down, or call the REST API of the RTVI-CV 2D detection / tracking microservice. Trigger when the user says things like 'deploy rtvi-cv', 'start warehouse 2d', 'add a stream', 'check rtvi-cv health', or 'stop the perception container'. Not for VLM, embedding, or analytics — use the matching vss-* skill.

nvidia avatarnvidia
获取
nemo-mbridge-perf-megatron-fsdp

nemo-mbridge-perf-megatron-fsdp

3Ktesting-qa

Operational guide for enabling Megatron FSDP in Megatron-Bridge, including config knobs, code anchors, pitfalls, and verification.

nvidia avatarnvidia
获取
vss-manage-alerts

vss-manage-alerts

3Kbackend-api

Use for VSS alert workflows — real-time monitoring, Alert-Bridge subscriptions, Slack notifications, incident queries, camera onboarding. Not for non-alert analytics.

nvidia avatarnvidia
获取
nemo-mbridge-perf-moe-dispatcher-selection

nemo-mbridge-perf-moe-dispatcher-selection

3Ktesting-qa

Choose the right MoE token dispatcher (`alltoall`, DeepEP, or HybridEP) for the hardware, EP degree, and optimization stage. Summarizes patterns from DSV3, Qwen3, Qwen3-Next, and VLM bring-up work.

nvidia avatarnvidia
获取
vss-query-analytics

vss-query-analytics

3Kdevops-cloud

Use this skill when reading video-analytics metrics, incidents, alerts, and sensor data via the VA-MCP server (port 9901). Not for live VLM or incident-range narrative reports.

nvidia avatarnvidia
获取
nemo-mbridge-perf-tp-dp-comm-overlap

nemo-mbridge-perf-tp-dp-comm-overlap

3Kagent-workflows

Operational guide for enabling TP, DP, and PP communication overlap in Megatron-Bridge, including config knobs, code anchors, pitfalls, and verification.

nvidia avatarnvidia
获取
mcore-linting-and-formatting

mcore-linting-and-formatting

3Kbackend-api

Linting and formatting for Megatron-LM. Covers running autoformat.sh, tools (ruff, black, isort, pylint, mypy), and code style rules.

nvidia avatarnvidia
获取
dicom-series-to-volume

dicom-series-to-volume

3Kagent-workflows

Used for converting one CT DICOM series folder to a HU NIfTI volume with affine evidence. Not for multi-frame DICOM or clinical use.

nvidia avatarnvidia
获取
nemo-rl-docs

nemo-rl-docs

3Kresearch-knowledge

Documentation conventions for NeMo-RL. Covers docs/index.md updates and docstring format. Do NOT use for: bug fixes, test fixes, dependency bumps, refactoring, CI/CD changes, performance tuning, or any task that does not involve writing or updating documentation.

nvidia avatarnvidia
获取
nemo-automodel-model-onboarding

nemo-automodel-model-onboarding

3Ktesting-qa

Guide for onboarding new model architectures into NeMo AutoModel, including architecture discovery, implementation patterns, registration, and validation.

nvidia avatarnvidia
获取
nemo-automodel-launcher-config

nemo-automodel-launcher-config

3Kdevops-cloud

Configure NeMo AutoModel job launches for interactive runs, Slurm clusters, and SkyPilot cloud execution.

nvidia avatarnvidia
获取
dynamo-router-starter

dynamo-router-starter

3Kdevops-cloud

Start or patch Dynamo router modes and run router endpoint smoke checks. Use for round-robin, KV-aware, least-loaded, or device-aware routing setup; use recipe-runner for recipe deployment and troubleshoot for failure diagnosis.

nvidia avatarnvidia
获取
dynamo-troubleshoot

dynamo-troubleshoot

3Kbackend-api

Diagnose failed or unhealthy Dynamo deployments. Use when pods, model-cache jobs, PVCs, workers, frontend/router health, endpoints, or benchmark jobs fail; use recipe-runner/router-starter before this for normal bring-up.

nvidia avatarnvidia
获取
dynamo-recipe-runner

dynamo-recipe-runner

3Kdevops-cloud

Select, validate, patch, and deploy existing NVIDIA Dynamo Kubernetes recipes. Use for model/backend/GPU/deployment-mode recipe bring-up; use router-starter for router-only mode work and troubleshoot for broken deployments.

nvidia avatarnvidia
获取
physicsnemo-discover

physicsnemo-discover

3Kresearch-knowledge

Official NVIDIA-authored guidance for navigating PhysicsNeMo — pick the model, datapipe, or example for a SciML/AI4Science task (surrogates, forecasting, downscaling, physics-informed, inverse, generative). Points at existing files via live repo search; never writes code. Do NOT use for installation or environment setup, training-loop or other code authoring/scaffolding, contributor/CI/packaging questions, repo-specific questions in physicsnemo-sym/-cfd/-curator, or general (non-physics) ML/PyTorch.

nvidia avatarnvidia
获取
nemo-mbridge-perf-memory-tuning

nemo-mbridge-perf-memory-tuning

3Ktesting-qa

Techniques for reducing peak GPU memory in Megatron Bridge — expandable segments, PEFT + SP input re-gather, parallelism resizing, activation recompute, CPU offloading constraints, and common OOM fixes.

nvidia avatarnvidia
获取
nemo-mbridge-perf-cuda-graphs

nemo-mbridge-perf-cuda-graphs

3Ktesting-qa

Validate and use CUDA graph capture in Megatron Bridge, including local full-iteration graphs and Transformer Engine scoped graphs for attention, MLP, and MoE modules.

nvidia avatarnvidia
获取
vss-ask-video

vss-ask-video

3Kagent-workflows

Use this skill to ask the VSS agent's video_understanding tool a fresh visual question about a recorded clip. Not for prior tool output, search hits, or metadata-answerable questions.

nvidia avatarnvidia
获取
dynamo-interconnect-check

dynamo-interconnect-check

3Kdevops-cloud

Validate that a Dynamo deployment's NIXL/UCX/NCCL interconnect is ready for disaggregated serving over RDMA/NVLink. Use after recipe-runner brings a deployment up (especially disagg/multi-node) to confirm the KV transport is correct; use troubleshoot for diagnosing already-failed pods.

nvidia avatarnvidia
获取
想按分类查看?试试 /category/writing-content.