YOUR RESEARCH WORKSPACE
What are we investigating?
Start with a task, a puzzling result, or a hypothesis.
Work
with your agent, then let a focused study run on its own.
Enter to send · Shift + Enter for a new line. Choose an installed agent above, or set up an API in Connections.
A place to start
Choose an example and make it yoursYOUR COMPUTE
Machines
Connect your machines. Run experiments in parallel.
The research agent stays here. Connected machines run your experiments.
Worker jobs
Detailed receipts
YOUR AGENT, YOUR MACHINE
API connections, if you need them.
Choose your agent and model inside each conversation. Installed CLI agents need no API setup here.
SIMJECTURE BENCHMARK RESULTS
Model leaderboard
Measured scientific coding tasks. Compare completion, time and API-equivalent cost.
Task leaderboard
Finish time ranking
Median verified finish time · shorter bars are better
Cost ranking
API-equivalent cost per attempt · shorter bars are better
Red marks unfinished runs. They have no verified finish-time rank; cost bars still show recorded spending, including labelled lower bounds. Configurations without inference are omitted.
Cost–time Pareto plot
Upper left is better. Dashed lines connect the same model's reasoning efforts.
Published API tariffs
Your runs & community contributions
Test any supported model with your installed agent or API connection. Local and imported runs appear in Community & local. Official results are maintained by Simjecture.
Prepare or grade a task
Your benchmark conversations
Latest numerical grade
YOUR WORKBENCH
Research tools
Use what’s installed. Add a solver when your investigation needs it.
Start with Python. NumPy, SciPy, pandas and plotting are included. External solvers are optional.
Installed 1
On this computer · includes detected local builds
Available to add 0
Install directly or prepare with your agent
Installation and readiness checks tell you a tool can run. Each investigation still validates its scientific method. For tools on SSH workers, use .
The agent can edit project files and run commands on this machine.
AUTONOMOUS RESEARCH
PREPARE IT TOGETHER
Let your agent shape the investigation.
Your conversation and files are the starting point. The agent fills the brief; you review it before launching.
YOUR AGENT’S PROPOSAL
When the study finishes, your interactive agent will explain its report in this conversation. Each study keeps its own results.