MLOps & Model Serving Bundle | Prompeteer.ai
Deployment, inference serving, model registries, and production model monitoring.
Included Skills (18)
- Modal Serverless GPU — Deploys and runs Python functions on cloud GPUs, enabling ML model deployment and inference without infrastructure management for developers.
- Cosmos Embedder Tool — This tool enables developers to fine-tune, evaluate, and perform inference with Cosmos-Embed1 models for advanced video-text retrieval and semantic video analysis tasks.
- ClearML Guidance — Provides expert assistance for ClearML, the open-source MLOps platform, enabling developers to streamline ML experiment tracking and lifecycle management.
- Brev GPU Orchestrator — This tool enables developers to manage NVIDIA TAO training and inference workloads by automating GPU instance deployment and job execution via the Brev CLI.
- SLURM Cluster Executor — This tool enables AI agents to execute TAO training and inference jobs on remote SLURM GPU clusters using SSH, Pyxis, and Enroot containers.
- Defect Image Generator — This tool orchestrates NVIDIA Cosmos AnomalyGen workflows to generate synthetic defect datasets and perform inference for industrial inspection tasks like PCBA and surface analysis.
- Mask2Former Segmentation Trainer — This tool enables developers to train, evaluate, and deploy high-quality universal image segmentation models using the NVIDIA TAO Mask2Former framework.
- Sparse4D Perception Trainer — This tool enables developers to train, evaluate, and deploy Sparse4D models for high-performance multi-camera temporal 3D object detection and tracking tasks.
- Physical AI Infrastructure Manager — This tool assists engineers in deploying, scaling, and hardening resilient NVIDIA AI infrastructure across MicroK8s and Azure AKS for synthetic data generation workflows.
- DeepStream Model Importer — This tool automates the integration of object detection models into NVIDIA DeepStream pipelines, streamlining model acquisition, TensorRT engine building, and performance benchmarking.
- Video Data Orchestrator — This tool manages end-to-end video data augmentation and auto-labeling workflows on OSMO, assisting engineers with preflight, job submission, monitoring, and output retrieval.
- Megatron Resiliency Manager — This tool enables and manages fault tolerance, straggler detection, and automatic restart features to ensure stable training for Megatron Bridge users.
- GPU Memory Optimizer — This tool helps machine learning engineers resolve GPU out-of-memory errors and optimize peak memory usage in Megatron Bridge training environments.
- VCN Gap Analyzer — This tool identifies weak classification samples for NVIDIA TAO VCN experiments to help engineers optimize decision thresholds and target data for augmentation.
- CUDA Graph Optimizer — This tool helps developers optimize training performance by configuring and validating CUDA graph capture implementations within Megatron Bridge for various model architectures.
- AOI Model Optimizer — Automates the full DEFT improvement loop for NVIDIA TAO PCB inspection models to achieve specific false accept rate and recall performance targets.
- VLM Gap Analyzer — This tool identifies false-positive and false-negative failure cases from VLM binary classification predictions to support root-cause analysis and model improvement workflows.
- LLM Observability Expert — Assists developers in tracing and monitoring LLM applications using Langtrace, an open-source observability platform built on OpenTelemetry.