MLOps & Model Serving Bundle | Prompeteer.ai

Deployment, inference serving, model registries, and production model monitoring.

Included Skills (18)

  1. Modal Serverless GPU — Deploys and runs Python functions on cloud GPUs, enabling ML model deployment and inference without infrastructure management for developers.
  2. Cosmos Embedder Tool — This tool enables developers to fine-tune, evaluate, and perform inference with Cosmos-Embed1 models for advanced video-text retrieval and semantic video analysis tasks.
  3. ClearML Guidance — Provides expert assistance for ClearML, the open-source MLOps platform, enabling developers to streamline ML experiment tracking and lifecycle management.
  4. Brev GPU Orchestrator — This tool enables developers to manage NVIDIA TAO training and inference workloads by automating GPU instance deployment and job execution via the Brev CLI.
  5. SLURM Cluster Executor — This tool enables AI agents to execute TAO training and inference jobs on remote SLURM GPU clusters using SSH, Pyxis, and Enroot containers.
  6. Defect Image Generator — This tool orchestrates NVIDIA Cosmos AnomalyGen workflows to generate synthetic defect datasets and perform inference for industrial inspection tasks like PCBA and surface analysis.
  7. Mask2Former Segmentation Trainer — This tool enables developers to train, evaluate, and deploy high-quality universal image segmentation models using the NVIDIA TAO Mask2Former framework.
  8. Sparse4D Perception Trainer — This tool enables developers to train, evaluate, and deploy Sparse4D models for high-performance multi-camera temporal 3D object detection and tracking tasks.
  9. Physical AI Infrastructure Manager — This tool assists engineers in deploying, scaling, and hardening resilient NVIDIA AI infrastructure across MicroK8s and Azure AKS for synthetic data generation workflows.
  10. DeepStream Model Importer — This tool automates the integration of object detection models into NVIDIA DeepStream pipelines, streamlining model acquisition, TensorRT engine building, and performance benchmarking.
  11. Video Data Orchestrator — This tool manages end-to-end video data augmentation and auto-labeling workflows on OSMO, assisting engineers with preflight, job submission, monitoring, and output retrieval.
  12. Megatron Resiliency Manager — This tool enables and manages fault tolerance, straggler detection, and automatic restart features to ensure stable training for Megatron Bridge users.
  13. GPU Memory Optimizer — This tool helps machine learning engineers resolve GPU out-of-memory errors and optimize peak memory usage in Megatron Bridge training environments.
  14. VCN Gap Analyzer — This tool identifies weak classification samples for NVIDIA TAO VCN experiments to help engineers optimize decision thresholds and target data for augmentation.
  15. CUDA Graph Optimizer — This tool helps developers optimize training performance by configuring and validating CUDA graph capture implementations within Megatron Bridge for various model architectures.
  16. AOI Model Optimizer — Automates the full DEFT improvement loop for NVIDIA TAO PCB inspection models to achieve specific false accept rate and recall performance targets.
  17. VLM Gap Analyzer — This tool identifies false-positive and false-negative failure cases from VLM binary classification predictions to support root-cause analysis and model improvement workflows.
  18. LLM Observability Expert — Assists developers in tracing and monitoring LLM applications using Langtrace, an open-source observability platform built on OpenTelemetry.