AI, machine learning and deep learning are nested ideas. Artificial intelligence is the broad goal of making machines do intelligent tasks. Machine learning is the main way to achieve it: models that learn from data. Deep learning is a powerful kind of machine learning that uses many-layered neural networks.

What is the difference between AI, ML and deep learning?

Artificial intelligenceMachine learningDeep learning
What it isAny technique that makes computers act intelligentlySystems that learn patterns from data instead of hand-written rulesMachine learning with many-layered neural networks
ScopeBroadestA subset of AIA subset of machine learning
Typical dataAnyTables of features (numbers, categories)Images, audio, video and text
Data neededVariesSmall to mediumUsually large
ExamplesChess engines, route planners, assistantsLoan risk scores, spam filters, price predictionFace unlock, speech recognition, ChatGPT and Claude

What is artificial intelligence?

AI is the goal, not a single technique. Early AI used hand-written rules ("if the patient has symptom A and B, suggest test C"). Rules work for narrow problems but break when the world is messy, which is why most modern AI is built with machine learning.

What is machine learning?

Instead of writing rules, you give a model many examples and let it learn the pattern. Show it thousands of past loan applications with outcomes, and it learns to estimate risk for new ones. Classical algorithms such as linear and logistic regression, decision trees, random forests and gradient boosting are the workhorses of business data.

What is deep learning?

Deep learning uses neural networks with many layers. Each layer learns slightly more abstract features: edges, then shapes, then objects in an image; or letters, then words, then meaning in text. It needs more data and computing power, but it dominates tasks involving images, speech and language. Transformers, the architecture behind today's large language models, are a deep learning design. We explain them in how LLMs like ChatGPT and Claude work.

Which should I learn first?

  1. Programming and maths foundations: Python, basic statistics and linear algebra.
  2. Classical machine learning: it teaches training, validation, overfitting and evaluation clearly.
  3. Deep learning: neural networks, backpropagation, CNNs and RNNs.
  4. Transformers and LLMs: attention, tokenisation, fine-tuning and retrieval-augmented generation.

This is exactly the order in Program Zero: Phase 6 (classical ML), Phase 7 (deep learning) and Phase 8 (NLP, transformers and building LLMs).

The bottom line

Think of three circles, one inside the other: AI contains machine learning, which contains deep learning. Learn them from the outside in, and each step will make sense.