Old version
May 12, 2026 05:44 UTC
03becf091e2454db0d8976bfced05081cc44247c63151117b5b5e4373f22ba84
CA-V-002484
New version
May 24, 2026 00:42 UTC
5588c46ebc049462c101ca519412a0d3e4b6765502325eb49a53219e564b757c
CA-V-002947
Share 𝕏 Share in Share
0 Sentences added
38 Sentences removed
1 Sentences modified
41 Sentences before
3 Sentences after
Added
Removed
Modified
BeforeAfter
2Agentic RAG Cohere on Azure Responsible Use Security Usage Policy Command A Technical Report Command R and Command R+ Model Card Cohere Labs Cohere Labs Acceptable Use Policy More Resources Cohere Toolkit Datasets Improve Cohere Docs Light On this page Safety Benchmarks Intended Use Cases Unintended and Prohibited Use Cases Usage Notes Model Toxicity and Bias Toxic Degeneration Reinforcing Historical Social Biases Technical Notes Language Limitations Sampling Parameters Prompt Engineering Potential for Misuse Responsible Use Command R and Command R+ Model Card Copy page This documentation aims to guide developers in using language models constructively and ethically.2Agentic RAG Cohere on Azure Responsible Use Security Usage Policy Command A Technical Report Command R and Command R+ Model Card Cohere Labs Cohere Labs Acceptable Use Policy More Resources Cohere Toolkit Datasets Improve Cohere Docs Light Responsible Use Command R and Command R+ Model Card Copy page
3To this end, we’ve included information below on how our Command R and Command R+ models perform on important safety benchmarks, the intended (and unintended) use cases they support, toxicity, and other technical specifications. [NOTE: This page was updated on October 31st, 2024.] Safety Benchmarks The safety of our Command R and Command R+ models has been evaluated on the BOLD (Biases in Open-ended Language Generation) dataset (Dhamala et al, 2021), which contains nearly 24,000 prompts testing for biases based on profession, gender, race, religion, and political ideology.Removed
4Overall, both models show a lack of bias, with generations that are very rarely toxic.Removed
5That said, there remain some differences in bias between the two, as measured by their respective sentiment and regard for “Gender” and “Religion” categories.Removed
6Command R+, the more powerful model, tends to display slightly less bias than Command R.Removed
7Below, we report differences in privileged vs. minoritised groups for gender, race, and religion.Removed
8Intended Use Cases Command R models are trained for sophisticated text generation—which can include natural text, summarization, code, and markdown—as well as to support complex Retrieval Augmented Generation (RAG) and tool-use tasks.Removed
9Command R models support 23 languages, including 10 languages that are key to global business (English, French, Spanish, Italian, German, Portuguese, Japanese, Korean, Chinese, Arabic).Removed
10While it has strong performance on these ten languages, the other 13 are lower-resource and less rigorously evaluated.Removed
11Unintended and Prohibited Use Cases We do not recommend using the Command R models on their own for decisions that could have a significant impact on individuals, including those related to access to financial services, employment, and housing.Removed
12Cohere’s Usage Guidelines and customer agreements contain details about prohibited use cases, like social scoring, inciting violence or harm, and misinformation or other political manipulation.Removed
13Usage Notes For general guidance on how to responsibly leverage the Cohere platform, we recommend you consult our Usage Guidelines page.Removed
14In the next few sections, we offer some model-specific usage notes.Removed
15Model Toxicity and Bias Language models learn the statistical relationships present in training datasets, which may include toxic language and historical biases along race, gender, sexual orientation, ability, language, cultural, and intersectional dimensions.Removed
16We recommend that developers be especially attuned to risks presented by toxic degeneration and the reinforcement of historical social biases.Removed
17Toxic Degeneration Models have been trained on a wide variety of text from many sources that contain toxic content (see Luccioni and Viviano, 2021).Removed
18As a result, models may generate toxic text.Removed
19This may include obscenities, sexually explicit content, and messages which mischaracterize or stereotype groups of people based on problematic historical biases perpetuated by internet communities (see Gehman et al., 2020 for more about toxic language model degeneration).Removed
20We have put safeguards in place to avoid generating harmful text, and while they are effective (see the “Safety Benchmarks” section above), it is still possible to encounter toxicity, especially over long conversations with multiple turns.Removed
21Reinforcing Historical Social Biases Language models capture problematic associations and stereotypes that are prominent on the internet and society at large.Removed
22They should not be used to make decisions about individuals or the groups they belong to.Removed
23For example, it can be dangerous to use Generation model outputs in CV ranking systems due to known biases (Nadeem et al., 2020).Removed
24Technical Notes Now, we’ll discuss some details of our underlying models that should be kept in mind.Removed
25Language Limitations This model is designed to excel at English, French, Spanish, Italian, German, Portuguese, Japanese, Korean, Chinese, and Arabic, and to generate in 13 other languages well.Removed
26It will sometimes respond in other languages, but the generations are unlikely to be reliable.Removed
27Sampling Parameters A model’s generation quality is highly dependent on its sampling parameters.Removed
28Please consult the documentation for details about each parameter and tune the values used for your application.Removed
29Parameters may require re-tuning upon a new model release.Removed
30Prompt Engineering Performance quality on generation tasks may increase when examples are provided as part of the system prompt.Removed
31See the documentation for examples on how to do this.Removed
32Potential for Misuse Here we describe potential concerns around misuse of the Command R models, drawing on the NAACL Ethics Review Questions.Removed
33By documenting adverse use cases, we aim to empower customers to prevent adversarial actors from leveraging customer applications for the following malicious ends.Removed
34The examples in this section are not comprehensive; they are meant to be more model-specific and tangible than those in the Usage Guidelines, and are only meant to illustrate our understanding of potential harms.Removed
35Each of these malicious use cases violates our Usage Guidelines and Terms of Use, and Cohere reserves the right to restrict API access at any time.Removed
36Astroturfing: Generated text used to provide the illusion of discourse or expression of opinion by members of the public, on social media or any other channel.Removed
37Generation of misinformation and other harmful content: The generation of news or other articles which manipulate public opinion, or any content which aims to incite hate or mischaracterize a group of people.Removed
38Human-outside-the-loop: The generation of text that could be used to make important decisions about people, without a human-in-the-loop.Removed
39Was this page helpful?Removed
40Yes No Edit this page Previous Cohere Labs Acceptable Use Policy Next Built withRemoved
Stay ahead of the changes

Watch this before it changes again

Follow unlimited companies, monitor the clauses that matter across every platform, and get the full institutional analysis on what each change obligates you to do.