Terminal-Bench-Science: Evaluating AI agents on scientific research workflows
1 sources1 storiesFirst seen 8/28/2026Score30Mixed Progress
Single Source
Bigness
30
Coverage
13
Recency
85
Engagement
20
Velocity
0
Confidence
49
Clipability
60
Polarization
0
Claims
0
Contradictions
0
Breakthrough
50
Sentiment Mix
Positive0%
Neutral100%
Negative0%
Expert Signals
matt_d
author • 1 mention
Hacker News
source • 1 mention
Related Events
Show HN: Derive – An open home for AI artifacts and workflows
Uncategorized • 8/27/2026
AI Engineer Notebooks – free, framework-free RAG/agents/evals on Colab
Uncategorized • 8/28/2026
AI Claude Now Controls Lab Machines - Anthropic Breakthrough - Briefs Finance
LLMs • 8/28/2026
Meta Internally Projected $10 Billion Annual Spend on Anthropic AI Models - marketscreener.com
LLMs • 8/27/2026
OpenAI, Anthropic Urge Cyber Defense Action as AI Models Improve - Bloomberg.com
LLMs • 8/27/2026
Timeline (1 stories)
Aug 28 12:58 AMFirst
Terminal-Bench-Science: Evaluating AI agents on scientific research workflowsHacker News124 engagement