Vibe Coding Discover

AI Agents

ScienceBuddy

View on GitHub

ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents

★ 10013 forksPythonMITGen-Verse

Research code and preview for ScienceBuddy, an interactive scientific agent workspace plus a double-recursive self-improvement experiment that alternates harness refinement with SkyRL GRPO model training on frozen scientific tasks.

Use Cases

Interpret scientific figures and organize related evidenceRetrieve literature and protein/database records for a research questionBuild evidence tables that separate retrieved records from gapsRun biomedical analyses (genomics, pharmacology, bioimaging) through 224 toolsInspect agent trajectories, tool inputs/outputs and artifactsImprove a Python agent harness via recursive self-improvementTrain a task model with GRPO driven by verifier scoresReproduce a 715/90/90 train/val/test scientific agent benchmark

Built With

Language
Python
Frameworks
SkyRL · vLLM · PyTorch · FlashAttention · pytest · uv · jsonschema

Tags

scientific-agents · interactive-agent · recursive-self-improvement · agent-harness · llm-agents · reinforcement-learning · grpo · agent-evaluation · biomedicine · tool-use · research-automation · trajectory-inspection · verifier · python

ScienceBuddy — Vibe Coding Discover