Skip to content
IT Canvass
Resources · Lesson

Datasets

Quick answer

Synthetic datasets of identities, accounts and entitlements you can load to practice aggregation, correlation and certifications.

Key takeaways

  • Synthetic identities and accounts
  • Safe, non-real data
  • Practice aggregation/correlation
  • Test roles and certifications

These synthetic datasets of identities, accounts and entitlements let you populate a sandbox with realistic-looking but entirely fake data, so you can practise the full governance loop, aggregation, correlation, roles, certifications, without touching real systems or real personal data.

What you get

  • Synthetic identity records
  • Sample accounts and entitlements
  • Data safe for training and testing
  • Enough volume to exercise real workflows

How to use them

Load a dataset as a source, aggregate it, correlate the accounts, then build roles and run a certification against it to experience the whole flow safely. Ideal for training and for testing customisations.

Common pitfalls

  • Mixing synthetic and real data in the same environment.
  • Assuming synthetic data matches your real data quality.
  • Using tiny datasets that hide performance behaviour.

Want to learn this properly?

Our live, instructor-led SailPoint Training covers this hands-on, with real projects and a certification path.

Check your understanding

  1. What are the datasets?

    • A. Aggregation, correlation, roles and certifications.
    • B. It lets you practice without exposing real data.
    • C. Synthetic identity/account data for safe practice.
    Show answer

    C. Synthetic identity/account data for safe practice.

    Synthetic identity/account data for safe practice.

  2. Why synthetic?

    • A. Synthetic identity/account data for safe practice.
    • B. It lets you practice without exposing real data.
    • C. Aggregation, correlation, roles and certifications.
    Show answer

    B. It lets you practice without exposing real data.

    It lets you practice without exposing real data.

  3. What can you practice?

    • A. Aggregation, correlation, roles and certifications.
    • B. Synthetic identity/account data for safe practice.
    • C. It lets you practice without exposing real data.
    Show answer

    A. Aggregation, correlation, roles and certifications.

    Aggregation, correlation, roles and certifications.

Frequently asked questions

What does the term Datasets refer to in SailPoint?

These synthetic datasets of identities, accounts and entitlements let you populate a sandbox with realistic-looking but entirely fake data, so you can practise the full governance loop, aggregation, correlation, roles, certifications, without touching real systems or real personal data.

What is the role of aggregation in Datasets?

Load a dataset as a source, aggregate it, correlate the accounts, then build roles and run a certification against it to experience the whole flow safely.

What tends to go wrong with Datasets?

Mixing synthetic and real data in the same environment. Assuming synthetic data matches your real data quality. Using tiny datasets that hide performance behaviour.
CallWhatsAppEnquire