AI UNIT 2 • STAGE 4 OF 5
Data doesn't choose itself. People do, and that's a responsibility
It's easy to talk about "the pile" like it just exists. But every pile of training data was assembled by people: people who decided what to include, what to leave out, and what counted as good enough. Those choices shaped the machine, and almost never included the communities being described.
Listen for the honest answer: the communities are rarely, if ever, asked.
Native scholars created a set of principles for exactly this, called CARE: data about Indigenous communities should benefit them, communities should have authority to control it, and those handling it hold a responsibility to the community, with ethics at the center.
Deciding whose examples represent your people isn't a small technical detail. It's a matter of sovereignty, of your nation's right to represent itself. You're now standing exactly where the data-sovereignty units later in this track begin.
In your Field Notes, write who you think should choose the training data about a community, whether communities were asked, and what fair choosing would look like.
Stage 5 brings the whole unit home with one question you're now ready to answer: when the examples are about your nation, who should choose them?
Take a position:
Saved automatically.