GitHub 项目简介: Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
README: Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture.
README: If you are interested in replicating something like ChatGPT out in the open
README: There is no trained model. This is just the ship and overall map. We still need millions of dollars of compute + data to sail to the correct point in high dimensional parameter sp…