eRightSoft All articles
Digital Equity

Who Gets to Build the Future? Open-Source AI and the Fight to Include Everyone

eRightSoft
Who Gets to Build the Future? Open-Source AI and the Fight to Include Everyone

Photo: diverse community members collaborating on computers data analysis, via 78.media.tumblr.com

Somewhere in America right now, a mortgage application is being denied, a résumé is being filtered out before any human eyes see it, and a benefits claim is being flagged for fraud—all by algorithms that neither the applicants nor their advocates are permitted to inspect. The systems making these determinations belong to corporations that treat their underlying logic as proprietary trade secrets. The people most affected by these decisions are, by design, kept in the dark.

This is the data divide. It is not simply a gap in internet access or computing power. It is a structural exclusion from the very mechanisms that increasingly govern economic and civic life in the United States. And it falls hardest on communities that already carry the heaviest burdens: low-income households, communities of color, rural populations, and first-generation immigrants navigating systems that were not designed with them in mind.

The Algorithm as Gatekeeper

The scale of algorithmic decision-making in American life is difficult to overstate. The Federal Trade Commission has documented how AI-driven tools influence credit scoring, insurance pricing, hiring, and child welfare assessments. The Department of Housing and Urban Development has raised alarms about discriminatory patterns in automated tenant screening. Academic researchers have repeatedly demonstrated that facial recognition systems perform significantly worse on darker-skinned faces, and that natural language processing models encode social biases absorbed from their training data.

Yet when affected individuals or advocacy organizations attempt to scrutinize these systems, they encounter a wall. Proprietary AI is, almost by definition, a black box. The companies deploying it are rarely required by law to explain their outputs, and in many cases they argue—successfully—that revealing their model architecture would constitute disclosure of trade secrets. The result is a paradox: the more consequential AI becomes, the less accountable it is to the people it affects most.

Open-Source as a Structural Correction

Open-source AI tools do not solve discrimination by themselves. But they remove a critical obstacle to accountability: opacity. When the code is visible, it can be scrutinized. When the training data is documented, it can be questioned. When the model is modifiable, communities can adapt it to their own needs rather than accepting whatever a vendor has shipped.

Frameworks such as Hugging Face's open model ecosystem, Google's TensorFlow (released under the Apache License), and Meta's LLaMA series have dramatically lowered the technical barrier to building and auditing machine learning systems. Organizations like the Alan Turing Institute and domestic nonprofits including Data for Black Lives and the Algorithmic Justice League have begun leveraging these tools to conduct independent audits of systems deployed in public-sector contexts.

For community organizations operating on shoestring budgets, the practical implications are significant. A civil rights legal clinic in Chicago, for example, does not need to license an enterprise AI platform to analyze patterns in eviction filings. A tribal nation in the Southwest does not need a Fortune 500 technology partner to build a language preservation tool trained on its own cultural data. Open-source infrastructure makes these projects possible without surrendering data sovereignty or editorial control to a third party.

The Capacity Gap and How to Close It

Acknowledging the potential of open-source AI is not the same as pretending the barriers are trivial. Technical literacy remains unevenly distributed, and the communities with the greatest stake in algorithmic accountability are often the least resourced to develop it independently. A dataset does not audit itself; a model does not emerge from nothing.

This is where investment in digital equity infrastructure becomes inseparable from the conversation about AI justice. Organizations such as Code for America, the Mozilla Foundation's Responsible AI program, and a growing network of community technology centers are working to bridge this gap—offering training, technical assistance, and shared infrastructure to groups that lack in-house data science capacity.

Federal policy has begun, haltingly, to respond. The Biden administration's Blueprint for an AI Bill of Rights articulated principles around explainability and non-discrimination, though it stopped short of enforceable mandates. The National Science Foundation has expanded its Convergence Accelerator program to fund AI projects with explicit community benefit components. These are meaningful steps, but they remain insufficient against the scale of commercial AI deployment.

What is needed, advocates argue, is a sustained public commitment: funding for community-controlled AI literacy programs, requirements for algorithmic impact assessments in publicly funded systems, and open data mandates that give civil society the raw material to conduct independent research.

Building From the Inside Out

Perhaps the most powerful argument for open-source AI is not defensive—not merely about auditing what already exists—but generative. Communities that develop the capacity to build their own tools are no longer waiting for technology to be done to them. They are participating in its creation.

The Māori Data Sovereignty Network in New Zealand has pioneered frameworks for indigenous communities to assert ownership over data about themselves. Domestic equivalents are emerging: the Detroit Digital Justice Coalition has worked to ensure that smart city initiatives serve residents rather than extract from them. In Atlanta, historically Black colleges and universities have begun partnering with open-source AI initiatives to develop research capacity that reflects the priorities of their communities.

These efforts share a common thread: they treat open-source not as a technical preference but as a political stance. When code is open, power is at least potentially distributed. When it is closed, power consolidates—and communities that were already marginalized find themselves further from the levers of a system that is making ever more decisions about their lives.

The Stakes of Inaction

The trajectory of AI development in the United States is not predetermined. Choices made now—by policymakers, by funders, by technologists, and by the communities most affected—will shape whether artificial intelligence becomes a tool of broader human flourishing or an engine of accelerated inequality.

Open-source AI is not a panacea. It requires sustained investment, genuine community partnership, and policy frameworks that enforce accountability regardless of whether a system's code is publicly visible. But it represents a necessary condition for the kind of participatory technology governance that a democratic society should demand.

The data divide is real. So is the possibility of closing it—if the will exists to do so.

All articles

Related Articles

Wired Out: How Broadband Monopolies Abandoned Rural America — and What Communities Are Building Instead

Medicine's Blind Spot: Why Open-Source AI May Be the Only Cure for Algorithmic Discrimination in Healthcare

The Free Classroom Nobody Is Using: How America Is Squandering Its Open Coding Education Revolution

The Free Classroom Nobody Is Using: How America Is Squandering Its Open Coding Education Revolution