Improving Model Safety Behavior with Rule-Based Rewards
We’ve developed and applied a new method leveraging Rule-Based Rewards (RBRs) that aligns models to behave safely without extensive human data collection.
Fonte
openai.com
1.156documentos
We’ve developed and applied a new method leveraging Rule-Based Rewards (RBRs) that aligns models to behave safely without extensive human data collection.
Compliance API integrations, SCIM, and GPT controls to support compliance programs, data security, and user access at scale
Discover how prover-verifier games improve the legibility of language model outputs, making AI solutions clearer, easier to verify, and more trustworthy for both humans and machines.
OpenAI and Los Alamos National Laboratory are working to develop safety evaluations to assess and measure biological capabilities and risks associated with frontier models.
CriticGPT, a model based on GPT-4, writes critiques of ChatGPT responses to help human trainers spot mistakes during RLHF
We’re partnering with TIME and its 101 years of archival content to enhance responses and provide links to stories on Time.com
Highlighting innovative research and AI integration in cybersecurity
Diffusion models have significantly advanced the fields of image, audio, and video generation, but they depend on an iterative sampling process that causes slow generation.
We present a holistic approach to building a robust and useful natural language classification system for real-world content moderation.
Consistency models are a nascent family of generative models that can sample high quality data in one step without the need for adversarial training.
Paf adopted ChatGPT Enterprise across its entire company, with engineers using custom GPTs on a daily basis to speed up routine development tasks. Paf also integrated ChatGPT Enterprise into the grit:lab coding academy (gritlab.ax), training the next generation of software developers using an AI-augmented, systems-architecture mindset from day one. In addition to the wide range of use cases for developers…
Color Health is working with OpenAI to pioneer a new way of accelerating cancer patients’ access to treatment. Their new Cancer Copilot application uses GPT-4o to identify missing diagnostics and create tailored workup plans, enabling healthcare providers to make evidence-based decisions about cancer screening and treatment.
Nakasone brings cybersecurity experience to growing Board of Directors; will join the Board’s Safety and Security Committee
OpenAI and Apple announce partnership to integrate ChatGPT into Apple experiences.
Exploring the technology behind our text-to-speech model.