comments (10)

  • Just realized that there are basically no American open models right now ever since the Llama series was abandoned. Basically Gemma and GPT-OSS I guess?

    Ah but Mira Murati's new Inkling is Apache 2.0

    But it makes sense that if you're a university researcher you are thinking about what's a model that will be open weight and developed over the long term and doesn't raise 'Chyna' concerns in Washington DC

    firasd

  • I'm interested to see where they want to land performance-wise (i.e. which point they choose on the scaling curve) and the niche they want to carve. They have a decent ways to scale beyond trinity large, in paticular on posttrain/RL before they are competitive with open-weights, especially internationally.

    Deepseek is explicitly banned [1] at LLNL and I wouldn't be suprised if there's a blanket ban on all Chinese models. But nowadays models like tera/luna could fill this area of the pareto front, and LANL already runs openai models on their clusters [2]. Maybe it's in custom SFT/RL, for instrument control or sensitive topics? But you'll still have to compete with frontier models + a harness.

    I would have also liked to see a carrot tied to their offer. It'll be hard to get teams to contribute RL gyms or curated text. But throw in a "we'll fund a postdoc/student to do that" and I think you'd have teams scrambling to apply.

    [1] https://hpc.llnl.gov/about-livermore-computing/ai-ml-lc/lc-l...

    [2] https://www.energy.gov/nnsa/articles/nnsas-los-alamos-nation...

    lithobraking

  • There's no mention of "LLM" nor "language". It does mention "foundation model" which includes LLMs but that also includes non-LLM architectures and non-text data. Many of the Genesis Initiative proposals answer "foundation model" call with non-LLM systems. All the FM's I know about currently in this sphere are non-LLMs. The "about gs1" page also does not mention "LLM" but does talk more about agentic harness and workflows. That description certainly sounds LLM'ish but describes a more rich system. I don't mean to suggest that LLMs will not be part of these "genesis open models" but as described, this will not result in a replacement for the "claude" or "codex" commands.

    frumiousirc

  • It would be extremely interesting to me if the usgov produces a model which honors copyright and is also useful. This would give them extreme leverage over the labs, who may be violating copyright in significant and obvious ways.

    sroerick

  • If you’re in the weird position of knowing more about the national labs than the AI lab scene (like I am), link to a TechCrunch profile:

    https://techcrunch.com/2026/01/28/tiny-startup-arcee-ai-buil...

    My question is: why is this being run out of Argonne? Why not NERSC proper?

    nxobject

  • Aeroi

  • What would the selected participants get from this? Looks like there is no offer of funding?

    Smith42

  • Do all these models have any significant architectural differences or training data sources? What are the factors going into the diversity of their performance?

    an0malous

  • Does Europe have an equivalent program?

    andsoitis

  • Contributing to a project like this seems like a great way to get yourself export controlled

    victor9000