foundation-models
SciReasoner, a science AI whose reasoning experts prefer 98% of the time News
SciReasoner, a multimodal scientific foundation model that turns molecular and material structures into a shared vocabulary, hit state-of-the-art on 67 of 86 benchmarks, and in blind review domain experts preferred its explanations over frontier LLMs in 98% of cases.
What Are Vision-Language-Action Models? Lesson
A vision-language-action (VLA) model is a single neural network that takes in camera images and a plain-language instruction and outputs the actual motor commands to carry it out, letting one model both understand a scene and physically act on it.
Orca proposes a single 'world latent space' to replace next-token, next-frame, and next-action prediction News
Researchers introduced Orca, a world foundation model that learns one unified latent space from multimodal signals and predicts the next world state rather than the next token or frame, outperforming similar-sized specialists on text, image, and action tasks.
Evo Tool
Arc Institute's family of genome language models, released openly with code and checkpoints. Used by Arc and Stanford to generate complete synthetic bacteriophage genomes that were then built and tested in the lab against non-pathogenic bacterial hosts.