arxiv:2609.19134
Published on Sep 16
· Submitted by
Ling Yang on Sep 17
Upvote
99
Authors:
,
,
,
,
,
,
,
,
,
,
,
,
,
,
,
,
,
,
Abstract
Scientific code repositories encode decades of human knowledge in executable models, methods, and tools. Yet fragmented toolchains, implicit domain conventions, and specialized correctness criteria make this knowledge difficult to convert into reliable learning experience-a challenge we call the scientific experience bottleneck. We introduce ScienceIDE, infrastructure for turning the world's scientific code into programmable environments for scientific agents. Guided by expert-defined scientific cases and acceptance criteria, agents transform repositories into executable environments that support task generation, execution, and scientific verification. These environments provide a shared foundation for supervised fine-tuning, reinforcement learning, and evaluation. Using verified interaction trajectories, we train PhAI-IDE-72B, PhAI-IDE-9B, and PhAI-IDE-4B. The model family shows gains in held-out scientific-code repair and across selected general-purpose benchmarks in code, reasoning, and knowledge, providing evidence of positive transfer from scientific experience to broader capabilities. ScienceIDE lays the foundation for an integrated workspace for agent learning and scientific practice, making humanity's scientific software a shared substrate for developing scientific intelligence. Code: https://github.com/aitofound/ScienceIDE
View arXiv page View PDF Project pageGitHub 82 Add to collection
Community
Paper submitter 6 days ago
Code: https://github.com/aitofound/ScienceIDE Models: https://hf-mirror.com/collections/AItonomy/scienceide-model-series
6 days ago
This comment has been hidden (marked as Spam)
5 days ago
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
-
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations (2026)
-
Apodex 1.1: Scaling Agentic Intelligence for Complex Work (2026)
-
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? (2026)
-
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work (2026)
-
Self-Evolving Coding Agents (2026)
-
FACET: Preserving Source Intent and Executable State in Terminal Task Synthesis (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on HF Mirror checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Upload images, audio, and videos by dragging in the text input, pasting, or clicking here.
Tap or paste here to upload images
· Sign up or log in to comment
Upvote
99
Get this paper in your agent:
hf papers read 2609.19134
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash
Models citing this paper 0
No model linking this paper
Cite arxiv.org/abs/2609.19134 in a model README.md to link it from this page.
Datasets citing this paper 0
No dataset linking this paper
Cite arxiv.org/abs/2609.19134 in a dataset README.md to link it from this page.
Spaces citing this paper 0
No Space linking this paper
Cite arxiv.org/abs/2609.19134 in a Space README.md to link it from this page.