Prime Intellect 当前平台已经把 RL Environment(强化学习环境)→ Evaluation(评估)→ Training(训练)→ Inference(推理)→ Agent 串成一个开放体系
评论