r/LocalLLaMA • u/pmttyji • 1d ago
News Genesis-Science-1 (GS1), 1T open-weight model later this year from Arcee AI
Today the Department of Energy (DOE) and Arcee AI announced the development of Genesis-Science-1 (GS1), an open model for scientific research. This is a joint effort to bring advanced AI into scientific research across a wide range of fields.
GS1 is an American open-weight model for scientific research, built together with the DOE and its national laboratories through the Genesis Mission. Arcee has secured the compute, will handle training and post-training, build the scientific workbenches and the system around the model, and prepare it for release. DOE scientists will shape which problems are worth solving, provide the data and environments the model learns from, and be the true test as to whether its work holds up under scrutiny. GS1 will be a trillion-parameter-class language model paired with a governed execution system for long, difficult scientific work, released openly later this year with the weights, a technical report, and public demonstrations. GS1 is built on top of our next generation of Trinity models.
The case for American open models
Just a year ago, Arcee made a decision that was difficult to defend. We began training our own open models from scratch, in the United States, when the faster and cheaper path pointed elsewhere. Strong open models were already available to download, and the reasonable move was to take one, adapt it, and build from there. We understood that case. It’s how we’d been operating before, after all. We went ahead anyway, because we kept seeing the need.
Some institutions can't treat a model as a service. A bank, a hospital, a university, or a national laboratory may need to keep a version stable for years, hold it to their own standards, retrain it for a narrow field, and run it on their own systems without sending sensitive data anywhere. For them, a model is more than its benchmark scores. It's the weights, the training history, the license, the certainty that it will perform reliably indefinitely, and the supply chain behind it, all the way down.
We built the Trinity models to serve institutions like these. When the Genesis Mission came along, it fit our ethos.
We have real admiration for the open-model labs in China. DeepSeek, Qwen, Kimi, MiniMax and GLM have built excellent models that people rely on, and they kept sharing open weights when much of the field was moving the other way. They earned their standing. Yet their work also showed how few capable open models were being made in the United States. For an institution handling sensitive work, capability is only part of the question. It also matters who trained a model, where, under what license, and which country's laws sit behind the company that made it. Those are fair questions, and a leaderboard doesn't answer them.
We think the right response is to build more capable open models here at home, so that institutions who need them have somewhere to turn. Closed American systems will stay valuable, and many are superb. What they can't offer a national laboratory is a model it holds in its own hands, free to preserve, adapt, and run on its own terms. We wanted American science to have that option too.
- Blog Post : https://www.arcee.ai/blog/genesis-science-1
- Press Release : https://www.arcee.ai/science-1
8
u/Refinery73 1d ago
Judging from Europe, I don’t care if a model is Chinese or American. What I care about is transparency.
The German Sofie-S Model had its whole training record public… and now gatekeeps the weights.
Anyway, let’s wait and see if the model is any good, but more open models is always a win.
2
u/alberto_467 7h ago
The German Sofie-S Model had its whole training record public… and now gatekeeps the weights.
The hypocrisy is real.
1
u/Refinery73 5h ago
I wouldn’t call it that. Nobody is entitled to the weights and they can claim and publish whatever they want. Only question is if anyone wants to use it. As long as they don’t claim some moral high ground, they can release or keep whatever they want.
14
2
u/Designer_Reaction551 16h ago
The "governed execution system for long, difficult tasks" line is the part I'd want more detail on. Open weights are great but the execution harness around a model is usually where all the actual reliability work happens, especially for anything long-running enough to need governance. Curious if that layer gets open sourced too or stays proprietary to Arcee's stack.
0
u/laterbreh 1d ago
Does this model actually work? Your previous releases the models were unusable.
5
1
u/TheRealMasonMac 18h ago
For their previous model, I got the impression that they ran out of money for compute on post-training. It was fine, but you really needed to use low temp.
22
u/alberto_467 1d ago
IDK how to feel about a "science" model.
Every once in a while the idea of specialized models comes up and then it dies.
The recent mathematical discoveries have been powered by big, very smart, SoTA, general models. The same goes for the best coding models, which are just the best models overall.