Atlas · GenAI 2026
Natural Language Processing & Computer Vision
34 skills · ontology graph below shows relations within this section.
What this domain covers
This edition groups 34 capabilities in Natural Language Processing & Computer Vision across 5 named categories. The inventory contains 22 concepts and 12 tools; 14 skills appeared in at least two of the three original research runs. The remaining entries stay visible with their lower coverage so a reader can distinguish taxonomy scope from research-system agreement.
Current category labels: Audio & Speech · Computer Vision · Multimodal Vision · NLP Foundations · Text Understanding
Frequent learning foundations
- NLP supports 2 mapped skills
- Linear Algebra supports 1 mapped skill
Skills in this section
Audio AI
Audio & Speech
Computer Vision
Computer Vision
Object Detection
Computer Vision
OpenCV
Computer Vision
Vision-Language Models
Multimodal Vision
NLP
NLP Foundations
Tokenization
NLP Foundations
Multilingual NLP
Text Understanding
Named Entity Recognition
Text Understanding
Semantic Search
Text Understanding
Audio Processing
Audio & Speech
ElevenLabs
Audio & Speech
Librosa
Audio & Speech
Speech Recognition
Audio & Speech
Text-to-Speech
Audio & Speech
Whisper
Audio & Speech
Detectron2
Computer Vision
Emotion Recognition
Computer Vision
Facial Recognition
Computer Vision
Image Classification
Computer Vision
Image Segmentation
Computer Vision
MMDetection
Computer Vision
MediaPipe
Computer Vision
Object Tracking
Computer Vision
YOLO
Computer Vision
dlib
Computer Vision
Gensim
NLP Foundations
NLTK
NLP Foundations
spaCy
NLP Foundations
Information Retrieval
Text Understanding
Intent Detection
Text Understanding
Natural Language Understanding (NLU)
Text Understanding
Summarization
Text Understanding
Text Classification
Text Understanding