di-agent-knowledge-engine-datastage
Installation
SKILL.md
DataStage Parallel Engine
When to Use DataStage
- Batch ETL processing of large data volumes
- Parallel processing across multiple nodes
- Complex transformations with high throughput requirements
- Integration with enterprise databases and file systems
- Data warehouse loading and CDC operations
Engine Characteristics
- Parallel processing: Divides data into partitions processed simultaneously
- Pipeline parallelism: Multiple stages process different data concurrently
- Scalable: Add nodes to increase throughput
- High performance: Optimized for large-scale data movement