Search icon
Arrow left icon
All Products
Best Sellers
New Releases
Books
Videos
Audiobooks
Learning Hub
Newsletters
Free Learning
Arrow right icon
Pentaho 3.2 Data Integration: Beginner's Guide

You're reading from  Pentaho 3.2 Data Integration: Beginner's Guide

Product type Book
Published in Apr 2010
Publisher Packt
ISBN-13 9781847199546
Pages 492 pages
Edition 1st Edition
Languages

Table of Contents (27) Chapters

Pentaho 3.2 Data Integration Beginner's Guide
Credits
Foreword
The Kettle Project
About the Author
About the Reviewers
Preface
1. Getting Started with Pentaho Data Integration 2. Getting Started with Transformations 3. Basic Data Manipulation 4. Controlling the Flow of Data 5. Transforming Your Data with JavaScript Code and the JavaScript Step 6. Transforming the Row Set 7. Validating Data and Handling Errors 8. Working with Databases 9. Performing Advanced Operations with Databases 10. Creating Basic Task Flows 11. Creating Advanced Transformations and Jobs 12. Developing and Implementing a Simple Datamart 13. Taking it Further Working with Repositories Pan and Kitchen: Launching Transformations and Jobs from the Command Line Quick Reference: Steps and Job Entries Spoon Shortcuts Introducing PDI 4 Features Pop Quiz Answers Index

Time for action – validating genres with a Regex Evaluation step


In this tutorial you will read the modified films file and validate the genres field.

  1. Create a new transformation.

  2. Read the modified films file just as you did in the previous tutorial.

  3. In the Content tab, check the Rownum in output? option and fill the Rownum fieldname with the text rownum.

  4. Do a preview. You should see this:

  5. After the Text file input step, add a Regex Evaluation step. You will find it under the Scripting category of steps.

  6. Under the Step settings box, select Genres as the Field to evaluate, and type genres_ok as the Result Fieldname.

  7. In the Regular expression textbox type [A-Za-z\s\-]*(\|[A-Za-z\s\-]*)* .

  8. Add the Filter rows step, an Add constants step, and two Text file output steps and link them as shown next:

  9. Edit the Add constants step.

  10. Add a String constant named err_code with value GEN_INV and a String constant named err_desc with value Invalid list of genres.

  11. Configure the Text file output step after the Add...

lock icon The rest of the chapter is locked
Register for a free Packt account to unlock a world of extra content!
A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.
Unlock this book and the full library FREE for 7 days
Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of
Renews at €14.99/month. Cancel anytime}