• Español Español Spanish es
  • English English English en
  • Français Français French fr
  • Shopping Cart Shopping Cart
    0Shopping Cart
SIXE
  • Services
    • Technical Support
      • IBM Power
      • Ubuntu Pro with SIXE UP
      • Kubernetes
      • Openstack
      • Red Hat (RHEL)
      • SUSE
      • SAP
      • Ceph
      • IBM Storage
      • Databases
        • DB2
        • Informix
        • MariaDB / PostgreSQL
    • Migrations
      • To AIX from Red Hat or z/OS
      • To Linux
      • To Ubuntu from Red Hat and SUSE
      • To Red Hat OpenShift From VMWare
      • To OpenStack from VMWare and HyperV
      • To Proxmox VE from VMWare
    • Consulting & Strategy
      • Technology radar service
      • Sustainable technology
      • Ceph
      • DRP
      • vCTO
  • Infra
    • Servers
      • IBM LinuxOne 5 servers
      • IBM Power11
      • Supermicro
      • Dell
      • Lenovo
    • Storage
      • Ceph
      • IBM Storage FlashSystems
      • IBM Storage Scale (GPFS)
      • IBM Storage Fusion
      • Backup
        • IBM Storage Defender: Data Resilience & Ransomware Protection
        • Bacula
        • Trilio
    • Workloads
      • Containers
      • Virtual machines
      • Private Cloud
      • Databases
      • IBM Software
  • Cybersecurity
    • IT Security
      • Monitoring and Response (SIEM & SOAR)
        • Wazuh
        • IBM QRadar
      • Information protection (IRM and DLP)
        • Sealpath
      • IBM security tools
        • IBM PowerSC
    • OT Security
      • Claroty
      • Industrial Device Integrity and Security with Tenable OT (Indegy ICS)
      • Industrial & IoT cybersecurity solution with Indegy ICS and QRadar
  • AI ✨
    • AI agents for companies
    • On-Premise AI Inference
    • Training
    • Automation service
    • Docling
    • Ceph for AI and HPC
    • WatsonX
  • Training
    • Linux | AIX | IBM i
      • Linux
      • AIX
      • IBM i | i2040
      • PowerVM
      • PowerHA
    • Iaas & PaaS ☁️
      • OpenShift
      • Docker + K8s
      • KVM and oVirt (RHEV)
      • VMWare ESXi & vSphere
      • Red Hat Ansible
      • Foreman
    • Infra & Data
      • OpenStack
        • Deployment & Administration
        • Maintenance & Upgrades
        • Production & VMware Migration
      • CEPH
        • Deployment and administration
        • Advanced
        • Expert (Production Engineering)
      • HPC
      • Informix
      • Storage
      • DB2
    • IBM Security
      • IBM QRadar
      • Cybersecurity analyst
      • IBM Guardium
    • Official training
      • Official IBM Trainings
      • Canonical Course Catalog
      • Tailor-made training
  • About us
  • Contact us!
  • Menu Menu
Canonical business partner

Find other courses

×

Browse by field of study

IBM
🖥️ Mainframe & z/OS
z/OS 46CICS 9RPG 2DB2 38
📊 Data & Analytics
IBM Planning Analytics 4Infosphere 32BigInsights 7Datacap 12
🤖 AI & Automation
Watson 6IBM AI 15IBM AI 15
☁️ Cloud & Infrastructure
Cloud Management 4SmartCloud 5AIX 18IBM i 15IBM Storage 7Linux 15
🔐 Security 5
📁 Content Management
FileNet 14Content Navigator 7Case Manager 9Case Foundation 12Enterprise Records 8
⚙️ Development & Middleware
Integration Bus 3API Connect 6Rational 13DataPower 5Operational Decision Manager 8IBM Decision Server 4
💼 Business & Processes
Business Process Manager 29Sterling 15Curam 12TRIRIGA 5Operations Analytics 3IBM Algo One 17
Canonical / Ubuntu
🛠️ System Administration 4🔄 Automation 2🛡️ Cybersecurity 2☁️ Cloud 2📦 Containers 1🤖 AI 2💻 Virtualization 2
HashiCorp
🏗️ Terraform 3🔒 Vault 1
AI & Automation
⚡ N8N 2⚡ Zapier 2

Search by technology

  • AI
    • N8N
    • Zapier
  • HashiCorp
    • Terraform
    • Vault
  • Official Canonical Ubuntu training
    • AI
    • Automation
    • Containers
    • Cybersecurity
    • OpenStack Cloud
    • System administration
    • Virtualization
  • Official IBM Training
    • AIX
    • API Connect
    • Aspera
    • BigInsights
    • BladeCenter
    • Blockchain
    • Business Process Manager
    • Case Foundation
    • Case Manager
    • CICS
    • Cloud Management
    • Content Navigator
    • Curam
    • Datacap
    • DataPower
    • DB2
    • Enterprise Records
    • FileNet
    • IBM AI
    • IBM Algo One
    • IBM Cognos Analytics Training | BI Reporting Courses
    • IBM Decision Server
    • IBM Disk Systems
    • IBM Flash Storage
    • IBM i
    • IBM Planning Analytics
    • IBM Power Systems
    • IBM PureData Systems
    • IBM Security
    • IBM Storage
    • IMS
    • Informix
    • Infosphere
    • Instana
    • Integration Bus
    • Linux
    • Linux on IBM Power Systems
    • Miscellaneous
    • Open Platform
    • OpenPages
    • Operational Decision Manager
    • Operations Analytics
    • PowerHA
    • Rational
    • RPG
    • SmartCloud
    • Spectrum Suite
    • SPSS
    • Sterling
    • Tivoli
    • TRIRIGA
    • Turbonomic
    • Watson
    • WebSphere
    • z/OS
  • Wazuh

IBM InfoSphere DataStage v11.5 – Advanced Data Processing

2.240,00€

This course is designed to introduce you to advanced parallel job data processing techniques in DataStage v11.5. In this course you will develop data techniques for processing different types of complex data resources including relational data, unstructured data (Excel spreadsheets), and XML data. In addition, you will learn advanced techniques for processing data, including techniques for masking data and techniques for validating data using data rules. Finally, you will learn techniques for updating data in a star schema data warehouse using the DataStage SCD (Slowly Changing Dimensions) stage. Even if you are not working with all of these specific types of data, you will benefit from this course by learning advanced DataStage job design techniques, techniques that go beyond those utilized in the DataStage Essentials course.

SKU: 309. Category: Infosphere
  • Description

Description

At SIXE we have been providing official IBM training around the world for over 12 years. Get the best training from our specialists in Europe. We have important discounts and offers for two or more students.

Course details

IBM course code: KM423GCategory: IBM Infosphere / DataStage
Delivery: Online & on-site**Course length in days: 2

Target audience

Experienced DataStage developers seeking training in more advanced DataStage job techniques and who seek techniques for working with complex types of data resources.

Desired Prerequisites:

DataStage Essentials course or equivalent.

Instructors

The great majority of the IBM courses we offer are taught directly by our engineers. This is the only way we can guarantee the highest quality. We complement all the training with our own materials and laboratories, based on our experience during the deployments, migrations and courses that we have carried out during all these years.

Added value

Our courses are deeply role oriented. To give an example, the needs for technology mastery are different for developer teams and for the people in charge of deploying and managing the underlying infrastructure. The level of previous experience is also important and we take it very seriously. That is why beyond (boring) commands and tasks, we focus on solving the problems that arise in the day to day of each team. Providing them with the knowledge, competencies and skills required for each project. In addition, our documentation is based on the latest version of each product.

Agenda and course syllabus

Unit 1 –Accessing databases

Topic 1:  Connector stage overview

• Use Connector stages to read from and write to relational tables

• Working with the Connector stage properties

Topic 2:  Connector stage functionality

• Before / After SQL

• Sparse lookups

• Optimize insert/update performance

Topic 3:  Error handling in Connector stages

• Reject links

• Reject conditions

Topic 4:  Multiple input links

• Designing jobs using Connector stages with multiple input links

• Ordering records across multiple input links

Topic 5:  File Connector stage

• Read and write data to Hadoop file systems

Demonstration 1: Handling database errors

Demonstration 2:  Parallel jobs with multiple Connector input links

Demonstration 3:  Using the File Connector stage to read and write HDFS files

Unit 2 – Processing unstructured data

Topic 1:  Using the Unstructured Data stage in DataStage jobs

• Extract data from an Excel spreadsheet

• Specify a data range for data extraction in an Unstructured Data stage

• Specify document properties for data extraction.

Demonstration 1:  Processing unstructured data

Unit 3 – Data masking

Topic 1:  Using the Data Masking stage in DataStage jobs

• Data masking techniques

• Data masking policies

• Applying policies for masquerading context-aware data types

• Applying policies for masquerading generic data types

• Repeatable replacement

• Using reference tables

• Creating custom reference tables

Demonstration 1: Data masking

Unit 4 – Using data rules

Topic 1:  Introduction to data rules

• Using the Data Rules Editor

• Selecting data rules

• Binding data rule variables

• Output link constraints

• Adding statistics and attributes to the output information

Topic 2:  Use the Data Rules stage to valid foreign key references in source data

Topic 3:  Create custom data rules

Demonstration 1:  Using data rules

Unit 5 – Processing XML data

Topic 1:  Introduction to the Hierarchical stage

• Hierarchical stage Assembly editor

• Use the Schema Library Manager to import and manage XML schemas

Topic 2:  Composing XML data

• Using the HJoin step to create parent-child relationships between input lists

• Using the Composer step

Topic 3:  Writing Hierarchical data to a relational table

Topic 4:  Using the Regroup step

Topic 5:  Consuming XML data

• Using the XML Parser step

• Propagating columns

Topic 6:  Transforming XML data

• Using the Aggregate step

• Using the Sort step

• Using the Switch step

• Using the H-Pivot step

Demonstration 1:  Importing XML schemas

Demonstration 2: Compose hierarchical data

Demonstration 3: Consume hierarchical data

Demonstration 4:  Transform hierarchical data

Unit 6:  Updating a star schema database

Topic 1:  Surrogate keys

• Design a job that creates and updates a surrogate key source key file from a dimension table

Topic 2:  Slowly Changing Dimensions (SCD) stage

• Star schema databases

• SCD stage Fast Path pages

• Specifying purpose codes

• Dimension update specification

• Design a job that processes a star schema database with Type 1 and Type 2 slowly changing dimensions

Demonstration 1: Build a parallel job that updates a star schema database with two dimensions

 

Do you need to adapt this syllabus to your needs? Are you interested in other courses?  Ask us without obligation.

Locations for on-site delivery

  • Austria: Vienna
  • Belgium: Brussels, Ghent
  • Denmark: Cophenhagen
  • Estonia: Tallinn
  • Finland: Helsinki
  • France: Paris, Marseille, Lyon
  • Germany: Berlin, Munich, Cologne, Hamburg
  • Greece: Athens, Thessaloniki
  • Italy: Rome
  • Louxemburg: Louxembourg (city)
  • Netherlands: Amsterdam
  • Norway: Oslo
  • Portugal: Lisbon, Braga, Porto, Coimbra
  • Slovakia: Bratislava
  • Slovenia: Bratislava
  • Spain: Madrid, Sevilla, Valencia, Barcelona, Bilbao, Málaga
  • Sweden: Stockholm
  • Turkey: Ankara
  • United Kingdom: London

Related products

  • IBM BigIntegrate for Data Engineers v11.5.0.2

    1.225,00€
    Add to cart Add to cart Nº de alumnos Show Details Show Details Show Details
  • IBM InfoSphere Information Governance Catalog v11.5.0.2: Building the Catalog

    1.225,00€
    Add to cart Add to cart Nº de alumnos Show Details Show Details Show Details
  • IBM InfoSphere Information Governance Catalog v11.5.0.2: Understanding Your Information Assets

    1.225,00€
    Add to cart Add to cart Nº de alumnos Show Details Show Details Show Details
  • IBM Identity Governance and Intelligence Foundations

    3.075,00€
    Add to cart Add to cart Nº de alumnos Show Details Show Details Show Details

Blog!

  • AI Act: what actually applies on 2 August 2026
  • Storage for AI and HPC: Ceph, Lustre, GPFS or DAOS?
  • IBM QRadar, G2 Leader in SIEM, UEBA, NTA and IR 2026
  • EDR beyond Windows: protect Linux, AIX & IBM i
  • db2 formacion oficial sixe ibm
    Your team is doing manually what Db2 12 already handles

Services

24/7 emergency support
Hardware & systems
Migrations
On-premise AI & HPC
Training
Cybersecurity

Resources
Latest news
Whitepapers
Become a SIXE Partner
Request a demo

Partners

  • Canonical
  • Docker
  • HCL
  • IBM
  • Kaspersky
  • Lenovo
  • Red Hat
  • Sealpath
  • SUSE
  • Tenable

Our mission

SIXE ensures that your infrastructures based on IBM®, Red Hat®, Canonical®, and SUSE® technologies remain stable, secure, and optimized for as long as you need them.

At SIXE, we are committed to sustainability. That’s why we help you design, implement, and maintain durable and efficient technological solutions.

SIXE green energy

+34 91 198 02 43 (EU) | +1 628 900 3024 (US) | WhatsApp | Mon–Fri 8:30–16:30 (GMT+1)
© 2026 - SIXE | Training, consulting, professional services and turnkey projects | IBM, Lenovo, Canonical, Red Hat, HCL, Sealpath & SUSE Authorized Business Partner. Company registered in the INCIBE cybersecurity catalogue.
HQ - Madrid | Barcelona | Paris | Bruxelles  ·  About us | Become a Partner | Contact
  • Link to LinkedIn Link to LinkedIn Link to LinkedIn
  • Link to X Link to X Link to X
  • Link to Facebook Link to Facebook Link to Facebook
  • cookie policy
  • Legal notice
  • Privacy Policy
Link to: IBM InfoSphere Advanced QualityStage V11.5 Link to: IBM InfoSphere Advanced QualityStage V11.5 IBM InfoSphere Advanced QualityStage V11.5Link to: IBM InfoSphere Information Server Administrative Tasks V11.5 Link to: IBM InfoSphere Information Server Administrative Tasks V11.5 IBM InfoSphere Information Server Administrative Tasks V11.5
Scroll to top Scroll to top Scroll to top
SIXE
SIXE
Manage cookie consent

At SIXE, we prioritize your comfort and security. These technologies help us optimize our services, personalize content and offer you solutions tailored to what you really need. Your privacy is important: you will always be in control of what you share with us.

Funcional Always active
El almacenamiento o acceso técnico es estrictamente necesario para el propósito legítimo de permitir el uso de un servicio específico explícitamente solicitado por el abonado o usuario, o con el único propósito de llevar a cabo la transmisión de una comunicación a través de una red de comunicaciones electrónicas.
Preferences
El almacenamiento o acceso técnico es necesario para la finalidad legítima de almacenar preferencias no solicitadas por el abonado o usuario.
Statistics
El almacenamiento o acceso técnico que es utilizado exclusivamente con fines estadísticos. El almacenamiento o acceso técnico que se utiliza exclusivamente con fines estadísticos anónimos. Sin un requerimiento, el cumplimiento voluntario por parte de tu Proveedor de servicios de Internet, o los registros adicionales de un tercero, la información almacenada o recuperada sólo para este propósito no se puede utilizar para identificarte.
Marketing
El almacenamiento o acceso técnico es necesario para crear perfiles de usuario para enviar publicidad, o para rastrear al usuario en una web o en varias web con fines de marketing similares.
  • Manage options
  • Manage services
  • Manage {vendor_count} vendors
  • Read more about these purposes
See preferences
  • {title}
  • {title}
  • {title}