How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
//: # (![Hugging Face Collection(https://img.shields.io/badge/Models-fcd022?style=for-the-badge&logo=huggingface&logoColor=000)]())
💫SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation
⚙️ Algorithm Flow • 📊 Results ✨ Getting Started • 🏋️ Training • 🔧 Usage • 📃 Evaluation 🎈 Citation • 🌻 Acknowledgement • 📧 Contact • 📈 Star History
This is the implementation of the following paper:
The main features of gpoesia/minimo are: Unsupervised Reward Methods.
Projects with overlapping indexed features include: insightllm/rl-without-gt — [//]: # ([![Hugging Face Collection](https://img.shields.io/badge/Models-fcd022?style=for-the-badge&logo=huggingfac… kelaxon/ssr-zero — 💫SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation. leaplabthu/absolute-zero-reasoner — ⚙️ Algorithm Flow • 📊 Results ✨ Getting Started • 🏋️ Training • 🔧 Usage • 📃 Evaluation 🎈 Citation • 🌻… lili-chen/self-questioning-lm — Self-Questioning Language Models. njunlp/trans0 — Trans0 aims to initialize a multilingual LLM as a translation agent via monolingual data. This is a public version… chengsong-huang/r-zero — Check out our paper or webpage for the details.