
ETL and Data Pipeline Engineer
Persona
Schema and data-shape contracts across ingestion and warehouse boundaries — pipelines where upstream schema changes don't silently corrupt data.
About
Most etl and data pipeline work breaks on exactly the cases teams skip until production: schema and data-shape contracts across ingestion and warehouse boundaries, and idempotency, replay behavior, and duplicate prevention in reprocessing.
What Pipeline does:
- schema and data-shape contracts across ingestion and warehouse boundaries
- idempotency, replay behavior, and duplicate prevention in reprocessing
- batch/stream ordering, watermark, and late-arrival handling assumptions
- null/default handling and type coercion that can silently corrupt meaning
- data quality controls (completeness, uniqueness, referential integrity)
- observability and lineage signals for fast failure diagnosis
- backfill and migration safety for existing downstream consumers
What you get:
SOUL.md— Pipeline's identity and working methodologyetl-engineer.md— the full persona instruction fileMEMORY.md— session-persistent context template
Install:
# 1. Place in your project
cp etl-engineer.md .claude/personas/PIPELINE.md
# 2. Add to CLAUDE.md
echo "## Active Persona\nPipeline handles etl and data pipeline work. See: .claude/personas/PIPELINE.md" >> CLAUDE.md
# 3. Call by name in Claude Code
# Pipeline, [your task here]
Pipeline maps the problem space before writing a line, validates success and failure paths, and reports residual risk honestly. Use it when production discipline matters more than speed.
Core Capabilities
- schema and data-shape contracts across ingestion and warehouse boundaries
- idempotency, replay behavior, and duplicate prevention in reprocessing
- batch/stream ordering, watermark, and late-arrival handling assumptions
- null/default handling and type coercion that can silently corrupt meaning
- data quality controls (completeness, uniqueness, referential integrity)
- observability and lineage signals for fast failure diagnosis
- backfill and migration safety for existing downstream consumers
Customer ratings
0 reviews
No ratings yet
- 5 star0
- 4 star0
- 3 star0
- 2 star0
- 1 star0
No reviews yet. Be the first buyer to share feedback.
Version History
This persona is actively maintained.
March 26, 2026
v1.0.0 — Initial release
One-time purchase
$49
By continuing, you agree to the Buyer Terms of Service.
Creator
iceboks
Creator
Software engineer building production AI tools. Skills and personas for engineering, DevOps, and executive leadership. Free skills that actually work. Paid personas with real decision frameworks and three-tier memory. Our agents include setup scripts and instructions on how to install. I'm always open to feed back for improvements or feature requests
View creator profile →Details
- Type
- Persona
- Category
- Engineering
- Price
- $49
- Version
- 1
- License
- One-time purchase
Works With
Works with OpenClaw, Claude Projects, Custom GPTs and other instruction-friendly AI tools.
Recommended Skills
Skills that complement this persona.
Moltline Humanizer — Local Edition (MCP Server)
Engineering
Voice-matching editor for AI drafts, run locally: scan drafts for measurable AI tells with cited evidence, fingerprint YOUR real writing voice, get rewrite briefs with nu
$29
Moltline Data Desk — Local Edition (MCP Server)
Engineering
Paste-your-data analytics MCP server you run locally: CSV profiler, A/B significance test, correlation, growth rates, funnel analysis, cohort retention, trend forecasts —
$19
Moltline Catalog Connector — MCP Server for OpenClaw
Engineering
Wire all 126 Moltline personas & skills into any MCP client — free to install, 126 free skills included
$19