> ## Content Index
> Fetch the complete content index at: https://www.testingcatalog.com/llms.txt
> Use this file to discover other available public pages before exploring further.

# OpenAI introduces o1-mini reinforcement fine-tuning alpha in Playground
- URL: https://www.testingcatalog.com/openai-introduces-o1-mini-reinforcement-fine-tuning-alpha-in-playground/
- Published: 2024-12-07T12:23:53.000Z
- Updated: 2024-12-07T12:23:53.000Z
- Author: Alexey Shabanov
- Tags: ChatGPT News, Latest AI News

OpenAI recently announced a new feature for reinforcement fine-tuning that will become available in the OpenAI Playground. This announcement was made during the [second day](https://x.com/OpenAI/status/1865136373491208674?ref=testingcatalog.com) of their “12 Days of OpenAI” event. The feature will allow users to teach models to reason within specific domains, enabling custom-tuned models to deliver significantly better results in specialized tasks. Initially, this capability will be offered to a limited set of alpha users, with broader availability expected in Q1\. Interested users can apply for access through an alpha sign-up form.

💡

You can apply for reinforcement fine-tuning Alpha access via [this form](https://openai.com/form/rft-research-program/?ref=testingcatalog.com)

When this feature is introduced, users will gain access to a fine-tuning method selector in the UI. This selector will include three options in addition to the currently available supervised fine-tuning: direct reference optimization and reinforcement fine-tuning. It is anticipated that direct reference optimization may eventually be launched as a standalone feature with its own announcement.

![OpenAI Playground](https://storage.ghost.io/c/2a/1b/2a1b1782-8506-4d7d-bf53-ad3fb52e2a0f/content/images/2024/12/screenshot-platform_openai_com-2024_12_06-19_59_18.png)

For reinforcement fine-tuning, users will have the ability to specify a grader schema to define how model responses should be evaluated. Alternatively, they can use a prompt to generate this schema automatically, making the process more intuitive and flexible.

> I noticed during the "12 Days of OpenAI: Day 2" livestream today that the OpenAI Platform sidebar has a new icon, possibly related to one of the upcoming announcements - "Custom Voices"  
>  
> \- "Create a voice below or using the OpenAI API"  
> \- "Create a voice sample of yourself by… [pic.twitter.com/c6ZGZpHBwr](https://t.co/c6ZGZpHBwr?ref=testingcatalog.com)
> 
> — Tibor Blaho (@btibor91) [December 6, 2024](https://twitter.com/btibor91/status/1865109134066274444?ref%5Fsrc=twsrc%5Etfw&ref=testingcatalog.com)

In addition to reinforcement fine-tuning, other features are also under development. One notable tool will allow users to clone their own voice. By reading a specific paragraph of text, users can enable OpenAI to create a voice clone capable of pronouncing any text in their own voice. This feature will likely be restricted to users aged 18 and older, marking a significant step in simplifying voice cloning technology.