OpenScience scores 75.7% on Terminal-Bench Science, outperforming Codex

iconKuCoinFlash
Share
AI summary iconSummary
OpenScience, a research agent from Synthetic Sciences, a startup in the Y Combinator 2026 Winter Batch, scored 75.7% on Terminal-Bench Science, outperforming Codex. The tool automates experiments, writes code, and integrates with scientific databases. Its Autoresearch feature enables iterative testing. On-chain data reflects growing interest in AI-driven research tools. Inflation data remains a key metric for investors monitoring macro trends.
ME AI News: Synthetic Sciences, a company in the Y Combinator 2026 Winter Batch (YC W26), has officially launched OpenScience, an open-source research agent. It can search for papers, process data, write code, and interact with scientific databases. The newly added Autoresearch feature enables the AI to autonomously run consecutive experiments. For example, when training a model, researchers only need to say, “Lower the validation loss.” OpenScience will first establish a baseline, then propose modifications on its own, execute experiments iteratively, retain improvements, revert setbacks, and decide the next steps based on prior results. Users can also set limits on the number of runs, total duration, and stopping conditions—or specify constraints like “Only modify the optimizer from now on.” OpenScience automates the iterative research process that previously required constant human oversight: proposing hypotheses, executing experiments, comparing metrics, discarding failed approaches, and advancing to the next round. Experiments can run locally or be delegated to remote GPUs, SSH servers, or Slurm/PBS clusters. Every round’s code, results, and decisions are logged for easy auditability. It supports direct integration with ChatGPT/Codex subscriptions, allows connection to personal API keys or local models, and offers access to the official pay-as-you-go Ace service. In internal tests, OpenScience completed 53 out of 70 Terminal-Bench Science tasks, achieving a score of 75.7%. By comparison, the publicly reported score for Codex + GPT-6 Astra is 68.1%. (Source: BlockBeats)
Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.