Seven optional mlops training skills had stale APIs, config paths, image locations, and requirement pins. Verified against upstream and corrected: - torchtitan: removed TOML train_configs paths (replaced upstream by config registry) - trl-fine-tuning: PPO removed from TRL 1.x -> GRPO/RLOO; SFTTrainer tokenizer= -> processing_class - flash-attention: torch.backends.cuda.sdp_kernel (deprecated) -> torch.nn.attention.sdpa_kernel; corrected false FA3/FP8-in-pip claim (FA2 only) - accelerate: DeepSpeedPlugin instance not raw dict; --config_file expects accelerate YAML; auto_wrap_policy -> transformer_based_wrap - saelens: v6 nested training config (sae=/logger=); from_pretrained tuple -> from_pretrained_with_cfg_and_sparsity - tensorrt-llm: Docker Hub image 404 -> NGC nvcr.io; rc pin -> GA; CUDA req updated - nemo-curator: pip extras renamed; repo moved to NVIDIA-NeMo/Curator; 1.x pipeline rewrite noted |
||
|---|---|---|
| .. | ||
| accelerate | ||
| chroma | ||
| clip | ||
| faiss | ||
| flash-attention | ||
| guidance | ||
| huggingface-tokenizers | ||
| inference/outlines | ||
| instructor | ||
| lambda-labs | ||
| llava | ||
| modal | ||
| models/segment-anything-model | ||
| nemo-curator | ||
| obliteratus | ||
| peft | ||
| pinecone | ||
| pytorch-fsdp | ||
| pytorch-lightning | ||
| qdrant | ||
| research | ||
| saelens | ||
| simpo | ||
| slime | ||
| stable-diffusion | ||
| tensorrt-llm | ||
| torchtitan | ||
| training | ||
| whisper | ||