Improving Model Safety Behavior with Rule-Based Rewards
July 24, 20242 views1 min read
We’ve developed and applied a new method leveraging Rule-Based Rewards (RBRs) that aligns models to behave safely without extensive human data collection.
Share:
Источник: OpenAI Blog
Related News
AI
Neural NetworksOpenAI Releases GPT-5
OpenAI released GPT-5, a new AI model with improved capabilities in text generation, mathematics, and programming.
4h ago66
AI
Neural NetworksCursor capitalizes on GitHub frustration, launches rival hosting platform
6h ago17
Neural NetworksRobin Williams’ Instagram account brought back to fight ‘AI abuse’
8h ago12