Data Engineering Bundle | Prompeteer.ai

Pipeline architecture, dbt models, schema design, lineage docs, and data-quality skills for data engineers.

Included Skills (52)

  1. Data Enrichment Optimizer — This skill helps RevOps and data engineering teams maximize data quality and minimize costs by optimizing provider selection and waterfall enrichment sequences.
  2. Soda Data Validator — Assists developers in writing YAML-based data quality checks using Soda, ensuring data freshness, completeness, and validity for reliable data pipelines.
  3. GPU Dataframe Expert — Provides expert guidance for developers accelerating pandas workflows using NVIDIA cuDF and dask-cuDF to achieve high-performance GPU-based data processing and ETL.
  4. BigQuery Data Analyzer — Analyze large datasets using Google BigQuery, enabling data professionals to run SQL queries and build machine learning models.
  5. Embedded Analytics Database — This skill helps data professionals use DuckDB, an embedded analytical database, for local data exploration and ETL pipelines.
  6. Vector Pipeline Assistant — Provides expert guidance for Vector, a high-performance observability data pipeline, helping developers manage logs, metrics, and traces efficiently.
  7. Real-Time Analytics Pipeline — This skill helps AI agents build real-time analytics pipelines, including event ingestion, ClickHouse storage, and dashboard wiring for live metrics.
  8. Kafka Event Streaming — Facilitates building event-driven systems using Apache Kafka for users needing message streaming and real-time data processing.
  9. DLT Pipeline Builder — Assists developers in building data pipelines using the open-source dlt Python library for data loading and ingestion.
  10. CSV Data Processor — This skill parses and generates CSV files, assisting developers with data import/export, ETL pipelines, and report generation tasks.
  11. Portable Analytics Assistant — Provides expert guidance for Ibis, enabling developers to write portable analytics code that runs on various SQL backends.
  12. DynamoDB Skill Builder — This skill helps developers design and build efficient NoSQL databases using Amazon DynamoDB for scalable applications.
  13. Analytics Quality Framework — This skill helps data engineers and analysts establish robust testing, automated monitoring, and incident response protocols for analytics models and dashboards.
  14. Kibana Dashboard Builder — Configure Kibana for Elasticsearch data visualization, enabling users to build dashboards and explore logs effectively.
  15. Fluentd Log Management — Configure Fluentd for collecting, filtering, and routing logs across distributed systems, assisting developers and DevOps engineers with log aggregation.
  16. Performing Cloud Asset Inventory With Cartography — Run Cartography to sync AWS, GCP, or Azure resources into a Neo4j graph database,
  17. Elasticsearch Search Skill — This skill configures and uses Elasticsearch for full-text search, custom analyzers, and index management, assisting users with search workloads.
  18. Redis Application Builder — Assists developers in building applications leveraging Redis for caching, messaging, and real-time data management, improving performance and scalability.
  19. Test Coverage Optimizer — This tool identifies untested code paths and generates targeted unit tests to help developers systematically increase their project's overall test coverage metrics.
  20. Performing AWS Account Enumeration With Scout Suite — Run the agentless, open-source ScoutSuite tool (via pip install and the `scout` CLI)
  21. Parallel Batch Processor — This tool enables developers to execute high-performance, parallel file operations and bulk data transformations across large directories with integrated progress tracking and error handling.
  22. PDF Data Parser — This skill parses PDFs into structured data, extracting text, tables, and images for AI applications like RAG and document processing.
  23. Cosmos DB Manager — Manage Azure Cosmos DB resources, enabling developers to build globally distributed, multi-model database applications efficiently.
  24. SQL Query Optimizer — Analyzes and optimizes SQL queries for improved performance, assisting developers and database administrators in enhancing database efficiency.
  25. Cloud SQL Manager — This skill assists developers and database administrators in provisioning and configuring managed MySQL, PostgreSQL, and SQL Server instances within Google Cloud environments.
  26. Axiom Expert Guide — Provides expert guidance for developers using Axiom, the serverless log management and analytics platform, to ingest, query, and visualize data.
  27. MotherDuck SQL Assistant — Provides expert guidance for MotherDuck, helping developers perform SQL analytics on cloud data and build hybrid local-cloud data pipelines.
  28. Apache Arrow Guide — Provides expert assistance for Apache Arrow, enabling developers to utilize its high-performance columnar format for data interchange and processing.
  29. Cube Analytics Assistant — Provides expert guidance for Cube, helping developers define data models, create metrics APIs, and build consistent analytics features.
  30. Cloud Array Storage — This skill enables developers to implement efficient, parallelized, and cloud-native N-dimensional array storage for large-scale scientific computing and data analysis pipelines.
  31. Data Version Control — DVC tracks large datasets and ML models, builds reproducible pipelines, and enables experiment tracking for data scientists and ML engineers.
  32. Grafana Dashboard Builder — This skill helps users build interactive Grafana dashboards and set up alerts for data visualization and monitoring.
  33. Orama Search Assistant — Provides expert guidance for developers using Orama, a fast full-text and vector search engine for various environments.
  34. SQL Database Helper — Assists developers by writing SQL queries, optimizing database performance, and generating migrations for various database systems and ORMs.
  35. Drizzle Database Studio — Visually explore and manage databases with Drizzle Studio, a lightweight GUI for browsing, querying, and editing data.
  36. Regex Debugging Assistant — This tool helps developers troubleshoot, explain, and optimize regular expressions through visual breakdowns, plain English interpretations, and automated test case validation.
  37. Firecrawl Web Scraper — Firecrawl converts websites into clean, structured data for LLMs, RAG pipelines, and data workflows, simplifying web scraping for developers.
  38. ElectricSQL Data Sync — ElectricSQL syncs Postgres data to client devices in real-time, enabling local-first applications with offline support and real-time updates.
  39. Schema Versioning Manager — Automates database schema versioning with migrations, rollbacks, and CI/CD integration, benefiting developers and database administrators.
  40. Baserow Database Builder — Build database-powered applications using Baserow, the open-source, no-code database, assisting users in creating spreadsheet databases and managing relational data.
  41. Web Scraping Framework — Crawlee builds reliable web scrapers and crawlers, helping developers extract data from websites at scale with ease.
  42. Real-Time Analytics — This skill helps users build real-time analytics APIs using Tinybird for event data ingestion and SQL querying.
  43. Drizzle ORM Expert — Assists developers using Drizzle ORM to write type-safe SQL queries and manage database migrations efficiently.
  44. PGlite Expert — Assists developers in embedding a lightweight WASM Postgres instance into client-side apps and serverless functions for local-first data solutions.
  45. Spreadsheet Data Merger — This tool enables users to consolidate multiple CSV or Excel files through intelligent column matching, automated data deduplication, and robust conflict resolution strategies.
  46. Pull Request Auditor — This tool assists developers by performing comprehensive security, performance, and quality audits on pull request diffs to ensure robust code standards.
  47. MongoDB Database Assistant — Helps developers design schemas, build pipelines, manage indexes, and operate MongoDB clusters for flexible data storage and retrieval.
  48. Convex Backend Assistant — Helps developers build real-time reactive backends using Convex, simplifying database creation and client synchronization for modern web applications.
  49. Parallel Data Processor — Deploys swarms of sub-agents to execute massive, independent data processing tasks, helping users efficiently manage large-scale document analysis and content generation workflows.
  50. Technical Documentation Specialist — This agent generates structured, clear technical documentation and guides to help developers and stakeholders effectively communicate complex information to diverse audiences.
  51. Data Contract Manager — This framework helps data engineers and analysts establish, enforce, and audit reliable BI data contracts to ensure high-quality reporting and accountability.
  52. Analytics Instrumentation Architect — This skill helps product and data teams design, govern, and maintain standardized event tracking plans for robust analytics pipelines.