The 8.3 Billion ‘People’ Who Don’t Exist: AI Creates Virtual Humans To Test What You’ll Buy

By eliminating costly real-world testing panels, the tool promises to transform product development

Harvard and MIT MatrAIx AI platform simulating virtual human profiles
Researchers at Harvard and MIT have built MatrAIx, an open-source system simulating 8.3 billion virtual humans to let companies run overnight consumer trials ChatGPT

Researchers from Harvard and MIT have unveiled an open-source platform that simulates 8.3 billion digital personalities, allowing companies to test products overnight. Powered by advanced language models, these virtual profiles, under the name MatrAIx, replicate complex human decision-making with striking accuracy.

However, unexpected variations between different AI engines threaten to produce highly skewed outcomes for businesses relying on the data.

Simulating Global Human Behaviour for Rapid Product Testing

Simulating human behaviour on a global scale, the system incorporates 1,290 distinct factors for every digital persona, ranging from demographic profiles and skill levels to shopping routines and precise traits such as 'how long they hesitate when prices rise'.

Designed for rapid evaluation, the network allows commercial entities and governments to run overnight trials for new software, services or policy shifts across a synthetic audience as vast as humanity itself.

To bring these vast synthetic crowds to life, the system runs on cutting-edge language models such as GPT-5.5 and Claude 4.8, which 'act out' each profile throughout the testing process. Researchers report an impressive 91.5 per cent accuracy rate when matching the behaviour of these digital agents to their individual background files.

How Businesses Can Run Instant Digital Consumer Trials

Commercial enterprises operating in software and retail may bypass lengthy, expensive consumer panels by using MatrAIx to evaluate products in just a few hours. The framework gives organisations four distinct digital environments for running experiments: survey questionnaires, interactive chatbot interfaces, web browser sessions and mobile applications.

Operating within this framework, digital personas can carry out more than 1,000 preset tasks spread over 25 key industries, such as financial services, medical care and digital security. As a result, commercial entities can model how users may respond to a system patch or evaluate the direct consequences of a price hike.

Severe AI Engine Biases Threaten to Distort Commercial Results

Notwithstanding this milestone, underlying language models introduce inherent biases into the platform. As the academic team at Harvard and MIT notes, selecting a specific AI engine heavily sways outcome data, as demonstrated when price-sensitivity trials on a webpage produced starkly contrasting responses depending on which system managed the simulation.

During product trials, 98.3 per cent of profiles operated by GPT-5.5 showed hesitation before making a purchase, whereas the identical cohort controlled by Claude Opus 4.8 registered just 27 per cent reluctance. This heavy dependence on the selected engine underscores that, while synthetic testing remains statistically uniform, it cannot fully replicate true human psychology.

Open-Source Release and the Future Risks for Global Marketing

In keeping with academic tradition, the team released the MatrAIx code under a free MIT licence, offering an initial testing dataset of one million personas rather than the full global population. Comprising roughly 600,000 anonymised real-world records alongside 400,000 synthetic profiles built through dependency graphs, the trial sample relies predominantly on actual human data.

Software developers and academic researchers can access the dataset on Hugging Face and review the repository on GitHub to launch their own large-scale trials.

Widespread adoption of this framework could surreptitiously rewrite the playbooks for commercial marketing, consumer design and behavioural research, especially across political campaigning. Positioned as a digital twin or a virtual voodoo doll of humanity, the platform sits open to the public, carrying vast potential alongside significant risk.


Frequently Asked Questions

  • What is MatrAIx?
    MatrAIx is an open-source platform that simulates digital personalities for rapid product testing.
  • How accurate is MatrAIx in simulating human behavior?
    MatrAIx has a reported accuracy rate of 91.5% in matching digital agent behavior to their background files.
  • What are the potential risks of using MatrAIx?
    The platform may introduce biases due to the underlying language models, affecting the accuracy of the results.