Regarding data sovereignty, a lot of people think that data is magical and that any unit that is collecting data is somehow giving positive data. It doesn't work that way. Data is deemed to be dirty or clean. You have to sift through to get the quality out of the data.
AI is not the software on your phone or on your laptop; it's the data centres. The data used and the compute used need to be the back end of what AI is modelled on. In order for you to model your AI, you need to collect data. One of the reasons a lot of these large language models today have issues like biases is that they scrape the Internet to collect the data. Let's be honest. The Internet isn't exactly the most prestigious and pristine element. It's a bit of a cesspool at times.
Data sovereignty means that Canada is generating terabytes of data in every sector in municipalities, provincially and federally, and we need to find a strategy and build a system where we own our own data. Having that good, high-quality data potentially gives us an edge over our competitors.
