Unit of competency Outline
Date retreived
23/07/2026 9:41 PM AWST
23/07/2026 9:41 PM AWST
Whilst all efforts are made to provide accurate and timely information from the relevant source/documentation, please be aware that the information supplied may not be the most current version. The accuracy of the detail has not been confirmed by the Department and therefore should not be relied upon without first confirming the contents.
Build natural language processing models and pipelines
Build natural language processing models and pipelines
Unit of competency
National Code
ICTAII503
ICTAII503
State Code
ODT32
ODT32
TGA Status
Current
Current
DTWD Status
Approved
Approved
State Implementation and Classification
Approved Date
25/05/2022
Field of Education
020103 - Programming
Original Release Date
25/05/2022
Nominal Hours
55
Description
This unit describes the skills and knowledge required to develop and apply routine natural language processing (NLP) pipelines to support the configuration of organisational computers to decode, interpret and manipulate human speech and text.The unit applies to individuals who work across a wide range of information and communications technology (ICT) roles, including support technicians, system administrators, programmers and cloud computing engineers.No licensing, legislative or certification requirements apply to this unit at the time of publication.
Notes
Elements and Performance Criteria
1. Prepare to create NLP process
- 1.1 Confirm work brief and tasks according to organisational policies and procedures
- 1.2 Establish requirements for morphological, lexical, syntactic and semantic analysis, disclosure integration and pragmatic analysis
- 1.3 Obtain human language data sources according to work brief
2. Extract, model and predict language tokens
- 2.1 Identify fields and blocks of language content
- 2.2 Source tokenizer according to organisational policies and procedures
- 2.3 Identify and mark sentence, phrase and paragraph boundaries
- 2.4 Normalise and tag acronyms in document according to work brief
- 2.5 Predict parts of speech for each generated token
3. Process texts and pipeline tokens
- 3.1 Lemmatise text in required document according to work brief
- 3.2 Identify and filter out stop words in document
- 3.3 Assign grammatical components with dependency parsing tags
- 3.4 Define relationships between parent words and other words in required document
- 3.5 Identify noun phrases according to work brief
- 3.6 Conduct named entity recognition (NER) and label nouns with real-world concepts
- 3.7 Create pipeline and run each token through NER tagging model
- 3.8 Run coreference algorithm on each token
4. Finalise NLP pipeline and required documentation
- 4.1 Review and apply security classifications and required security level in consultation with required personnel
- 4.2 Secure and save NLP pipeline according to organisational policies and procedures
- 4.3 Seek and implement feedback on processes from required personnel
No information
No information
No information
| State Code | National Code | Title | Type |
|---|---|---|---|
| BGJ4 | ICT50220 | Diploma of Information Technology | Qualification |
| BFF9 | ICT40120 | Certificate IV in Information Technology | Qualification |