• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Report

OpenScience launches as an open-source agent, with claimed benchmark leads over Codex and Claude Code

A post describes Synthetic Sciences’ OpenScience as a free, Apache 2.0-licensed workbench for any provider’s model. It reports that OpenScience solved 53 of 70 Terminal-Bench-Science tasks, scoring 75.7%.

Y CombinatorYC
4 Sources, 13d ago, first seen 13d ago

TLDR

A post says Synthetic Sciences launched OpenScience as an open-source agent. It reports a 75.7% score on Terminal-Bench-Science, compared with 68.1% for Codex with GPT-6 Astra on a September 23 mirror of the public leaderboard. On Terminal-Bench 4.0’s 14 science tasks, the post reports 71.4% for OpenScience and 60.0% for Claude Code on Claude Fable 5.1.

Combined views

1.6K

4 Sources, first seen 13d ago

98 reposts

Combined views

1.6K

4 Sources, first seen 13d ago

98 reposts

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Featured Source

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Related

Eight hidden-state numbers are claimed to predict 68 later measurements in Hermes 8B

A post says the numbers came from one layer of a frozen Hermes 8B and predicted measurements several layers later, with error about 69–76% lower than a simple average baseline.

Commentator Urges Safety Testing For All Capable AI Models Regardless Of License

4 Sources

Rohan Paul@rohanpaul_aiSynthetic Sciences just launched its open-source OpenScience agent now outscores Codex and Claude Code on agentic science benchmarks. A free Apache 2.0 workbench for any provider's model. On Terminal-Bench-Science, 70 research workflows hosted by Stanford and the Laude Institute, OpenScience solved 53 tasks for 75.7%. Codex with GPT-6 Astra scored 68.1%, the top entry on a September 23 mirror of the public leaderboard. The gap widened on Terminal-Bench 4.0's 14 science tasks, where OpenScience scored 71.4% and Claude Code on Claude Fable 5.1 reached 60.0%. GitHub link in comment13d
Y Combinator@ycombinatorRT @SynScience: OpenScience is now the #1 scientific agent. Today it's out of beta and live on Product Hunt, with: • A new IDE: a faster,…12d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    OpenScienceRohan Paul

    4 Sources

    Rohan Paul@rohanpaul_aiSynthetic Sciences just launched its open-source OpenScience agent now outscores Codex and Claude Code on agentic science benchmarks. A free Apache 2.0 workbench for any provider's model. On Terminal-Bench-Science, 70 research workflows hosted by Stanford and the Laude Institute, OpenScience solved 53 tasks for 75.7%. Codex with GPT-6 Astra scored 68.1%, the top entry on a September 23 mirror of the public leaderboard. The gap widened on Terminal-Bench 4.0's 14 science tasks, where OpenScience scored 71.4% and Claude Code on Claude Fable 5.1 reached 60.0%. GitHub link in comment13d
    Y Combinator@ycombinatorRT @SynScience: OpenScience is now the #1 scientific agent. Today it's out of beta and live on Product Hunt, with: • A new IDE: a faster,…12d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet