Literature · wanglab
RL-tuned variant of BioReason-Pro — reinforcement-learning fine-tuning over BioReason’s SFT base for sharper biological reasoning across KEGG pathways and variant data.
1–30s
Runs on On-demand GPU
Runs on our servers with on-demand compute. A first run needs time to load the model; active capacity can be reused and scales down when idle.Run this model
About this model
RL-tuned variant of BioReason-Pro — reinforcement-learning fine-tuning over BioReason’s SFT base for sharper biological reasoning across KEGG pathways and variant data. Served through the generic Modal runner, which pulls wanglab/bioreason-pro-rl from the Hugging Face Hub and runs it as a text-generation pipeline. Inputs and outputs were derived from the repository's declared pipeline tag rather than written by hand — check the model card before relying on a result.
Standardized I/O contract
Every model in the Hub speaks the same contract, which is what lets the Router and the agent call any of them without special-casing.
Inputs
Prompt
Outputs
The model's completion.