Lesson 1

Data vs Information

Data is raw, unprocessed, and often meaningless on its own. Information is what happens when we give data context. Can you tell the difference?

Learning intention Understand the difference between data and information, and why context transforms one into the other.

The key distinction

Data is raw, unprocessed values. It can look completely random and useless until it's organised.

Information is what we get when data is processed, organised, structured, or presented in context so that it becomes meaningful.

Example

75, 92, 34, 12, 72, 100, 65

Without context, these numbers mean nothing.

With context: Maths test scores → Fail: 12, 34 · Pass: 65, 72, 75, 92, 100

Interactive: guess the data

Below, raw numbers will be revealed progressively. Guess what the data represents before the context clues make it obvious.

Round 1 of 5
Context so far
—

Discussion questions

  • Q

    Can you think of an example where the same data could be interpreted as two different types of information depending on context?

  • Q

    Why do databases need to store data in a structured way — what would happen if they didn't?

  • Q

    When TikTok tracks that you watched a video for 12 seconds vs 0.5 seconds — is that data or information? What information can they derive from it?

Homework Task

Is my phone actually listening to me?

You've probably experienced this: you were talking to a friend about something — say, hiking boots — and then ads for hiking boots appeared on your phone. But your microphone was never intentionally activated. How?

Task Using what you've learned about databases, data, and information: write a 300–400 word response explaining how technology companies can predict your interests without ever listening to your conversations.

What we know from this unit

Every action creates a database record

Every time you open an app, scroll past a post, pause on a video, tap a link, search for something, or even move your phone — a record is created in a database. Not just one record. Millions. Timestamped, geolocated, and linked to your unique profile ID.

Data becomes information at scale

Individually, "user paused video at 2:14" is meaningless data. But when 50 million people pause at the same moment in the same type of video, that becomes information: viewers disengage with a certain type of content. Display less of this in the future so that we can retain their attention and engagement more with our platform.

Behavioural correlation, not mind-reading

Your friend probably also has a phone. They may have searched for hiking boots already. Your database records show that you spend time with people who live near each other (location data), are the same age, and share similar interests — because you like the same posts. The algorithm doesn't need to listen to your conversation: it already knows you and your friend share a profile cluster.

Prompts to get you thinking

  • ?

    What kinds of data does your phone collect that you might not have thought about? (Think beyond what you actively post.)

  • ?

    How does the concept of a junction table relate to how social media links you to your friends' interests?

  • ?

    If a company has 3 billion users and each generates 100 data points per day — that's 300 billion rows of data per day. How important is database design and efficiency at that scale?

  • ?

    Is this a good thing or a bad thing? Where is the line between personalisation and surveillance?

Your response should address

  1. What types of data are being collected (passive vs active)
  2. How raw data becomes useful information (the transformation we studied in Lesson 1)
  3. The role of relational databases in linking data about you, your friends, and your behaviour
  4. Your own reflection: should people know more about this? Should there be limits?
🎯

Extension: Research the term "data broker" — companies that buy and sell personal data. How might your database skills help you understand what they're doing with that data?

Fun fact

Your phone actually IS listening... sort of.

Voice assistants (Siri, Google Assistant) do listen for wake words — but that's not what causes the ad coincidences most people experience. Studies have repeatedly failed to find evidence that apps secretly record conversations for ad targeting. The truth — massive correlated databases linking your behaviour to millions of similar users — is actually more impressive and more unsettling than simple eavesdropping.

Lesson 1

What is a Database?

LEARNING INTENTION: Develop an understanding of:

  • What a Database is
  • How it is different from a spreadsheet
  • How data is different from information

 

A database is the name given to a system that has a collection of organized data. At its most basic, a database can be very similar to a spreadsheet by containing a single table of data that can be used to calculate, chart and predict. When carefully designed, however, databases offer many advantages that are difficult to achieve in a spreadsheet, such as relational links, unnecessary duplication of data, and multi-user write-access. You will learn more about these as you work through the unit.

It’s important to understand the differences between data and information:

  • Data: The raw, unorganised facts that need to be processed. Data can be something simple and appear to be random and useless until it is organised.
    • 75, 92, 34, 12, 72, 100, 65
  • Information: When data is processed, organised, structured or presented in a given context so as to make it useful, it is called information.
    • Fail: 12, 34
    • Pass: 65, 72, 75, 92, 100
    • %
    • Maths Test

To do:

  1. Create a new folder within your Digital Tech folder on your OneDrive.

  2. Open up Microsoft Access. On the first page that appears, select “Blank desktop database”.
    figure1.1.jpg

  3. Change the File Name to “Netflix Subscribers”, select the folder icon to navigate to the folder you made in Step 1, then select [Create].
    figure1.2.jpg

  4. Congratulations, you’ve made your first database! Now, before you begin to run around the room high-fiving everyone, you’ll need to begin lesson 2, develop your database structure and input some data. You can high-five everyone after you’ve completed it! 

 

SUCCESS CRITERIA: 

  • You've setup your first database and it has been saved on a cloud drive (OneDrive)
  • Understand which of the following is data, and which is information:
    • 28
    • 13% fuel remaining
    • 47cm
    • 2.3, 4.7, 5.6