1
Process-Reward Tactic Evolution for Long-Horizon Bioinformatics Workflows
让AI在长流程生物信息学任务中自我优化,过程奖励策略进化出更稳的决策路径。
arXiv:2606.20839v1 Announce Type: new Abstract: LLM agents can write code and call tools, but reliable bioinformatics work requires long-horizon inter…