agentsclimarketplace

Feature engineering

Skill sairam0424/MindForge/.mindforge/skills/feature-engineering

MindForge: The Enterprise Agentic Framework for Claude Code & Antigravity. High-performance autonomous execution, wave-parallelism, and multi-tier governance for production-grade AI engineering.From the repository description

Install
npx -y skills add sairam0424/MindForge --skill feature-engineering

Assembled from the repository path, not quoted from the project. Check it against their README if it does not work.

One thing to look at

  • 0 stars0 stars. Stars are a popularity signal and not a quality one, but at this level it is likely that nobody has read this closely except its author, and you would be relying on your own review.

SKILL.md

3.3 KB, 489 tokens by cl100k_base, as published. Nobody here has run it

Skill — Feature Engineering

When this skill activates

This skill activates when building ML pipelines that require feature creation, transformation, or selection. Use when designing feature stores, implementing automated feature discovery, or optimizing model input representation.

Mandatory actions when this skill is active

Before writing any code

  1. Conduct exploratory data analysis to understand feature distributions, missing patterns, correlations, and domain-specific relationships
  2. Define feature engineering strategy: target encoding risks, temporal leakage prevention, train-test split boundaries, and cross-validation approach
  3. Document business logic for derived features with domain expert validation and interpretability requirements
  4. Establish feature quality metrics: null rates, cardinality, stability over time, and correlation with target variable

During implementation

  • Implement feature transformations within sklearn Pipelines or similar frameworks to prevent train-test leakage
  • Use robust scaling methods appropriate to distribution (StandardScaler for normal, RobustScaler for outliers, quantile for non-parametric)
  • Create temporal features with proper lag handling: rolling windows, exponential smoothing, seasonal decomposition, time-since-event
  • Encode categorical variables with strategy matching cardinality (one-hot <10 categories, target encoding >50, embeddings for high-cardinality)
  • Generate interaction features guided by domain knowledge and feature importance: polynomial, ratio, difference, product features
  • Handle missing values explicitly with strategy documented: imputation (mean/median/mode), indicator variables, or model-based imputation
  • Validate feature importance using multiple methods: permutation importance, SHAP values, and univariate tests to identify top contributors

After implementation

  • Create feature documentation with schema definitions, transformation logic, expected ranges, and update frequency
  • Build feature monitoring dashboards tracking distribution drift, missing rate changes, and correlation stability over time
  • Generate feature store integration with versioning, metadata tracking, and point-in-time correctness for temporal joins
  • Validate feature pipeline performance: transformation latency, memory usage, and batch vs online serving consistency

Self-check before task completion

  • All features are computed within transformation pipelines to prevent train-test leakage
  • Feature importance analysis identifies top 20 contributors with interpretable business meaning
  • Temporal features respect time boundaries and use only historically available information
  • Feature documentation includes transformation logic, expected distributions, and monitoring thresholds
  • Feature validation tests confirm stability across different time periods and data segments

What ships with it

Read from the repository

Just SKILL.md. No reference files, no scripts.

Keep looking

Skills are one crate of 325,949. Ordering is by how many stacks a row turns up in, so the top of any crate is what has actually been picked rather than what has the most stars.