Tether’s technology arm, Tether Data, has launched a new AI initiative under its research division QVAC, releasing what it describes as the world’s largest synthetic STEM-focused AI dataset alongside a privacy-focused local AI application. The dataset, called QVAC Genesis I, contains about 41 billion text tokens of fully synthetic, education-focused data designed to train and improve language models in core STEM disciplines, while the new app, QVAC Workbench, is positioned as a comprehensive local AI workspace that runs models on users’ own devices rather than in centralized clouds. According to Tether Data, QVAC Genesis I is built from open-source seeds and then expanded into a synthetic corpus specifically tuned for educational and scientific use cases, with coverage across subjects such as mathematics, physics, biology, and medicine, and validated against multiple academic and reasoning benchmarks. The goal is to “level the playing field” for AI research by making high-quality training data openly available to researchers, independent developers, and institutions outside Big Tech, under a Creative Commons Attribution–NonCommercial (CC-BY-NC 4.0) license. This move is part of a broader push by Tether Data and QVAC to decentralize AI development and reduce dependence on proprietary, closed datasets by building a public ecosystem of synthetic educational data that has since been expanded further with QVAC Genesis II to a combined 148 billion tokens across 19 domains. On the application side, QVAC Workbench is released as Tether Data’s first consumer-facing AI product, presented as a “local-first” environment for running and managing AI models directly on user hardware. The app is intended to demonstrate that advanced AI workflows—such as coding assistance, document analysis, and STEM problem-solving—can be performed without relying on centralized servers, thereby aligning with Tether Data’s stated aim of decentralizing intelligence and enhancing data privacy and user control. Together, QVAC Genesis I and QVAC Workbench are framed by Tether Data as an attempt to challenge the concentration of AI capabilities and training data within a small number of large technology companies by providing open, non-commercial datasets and tools oriented toward community-driven AI development. "entities":["Tether","Tether Data","QVAC","QVAC Genesis I","QVAC Genesis II","QVAC Workbench","Hugging Face"]}`

AI-generated background, compiled from web sources — not editorial content.

More coverage

Explore the topic

More on App

Comments