一键重装系统工具 | U盘启动盘制作工具 | 误删文件恢复软件 | 硬盘数据抢救专家 | 电脑蓝屏修复助手 | C盘空间清理神器 | 电脑驱动离线安装工具 | 微信聊天记录恢复工具 | 照片误格式化恢复 | 电脑密码破解清除工具 | 系统崩溃紧急救援盘 | 电脑加速优化大师 | 电脑开不了机怎么重装系统 | 回收站清空了怎么恢复 | 硬盘分区丢失数据恢复 | 电脑卡顿重装系统有用吗 | U盘插入提示格式化数据恢复 | 电脑中毒文件被隐藏恢复 | 忘记电脑开机密码怎么办 | 新硬盘分区对齐工具 | 旧电脑装Win10流畅工具 | SD卡照片删除恢复免费版 | 移动硬盘打不开提示损坏修复 | 电脑无故重启系统修复工具 | 电脑小白一键重装神器 | 程序员电脑环境配置助手 | 设计师电脑字体/素材恢复工具 | 网吧网管系统维护工具箱 | 财务人员电脑发票备份恢复 | 学生党免费电脑系统安装包 | 电脑维修师傅必备工具盘 | 游戏玩家电脑性能优化助手 | 办公白领误删文档恢复软件 | 自媒体视频素材恢复工具 | 网课录制视频损坏修复工具 | 最好的U盘PE系统排名 | 数据恢复软件哪个最强 | 免费电脑助手与收费版区别 | 国产装机工具哪款无广告 | 离线版驱动助手推荐 | 轻量级电脑优化工具对比 | 支持NVMe驱动的PE工具 | 带网络功能的应急启动盘 | 2026最新版万能装机工具 | 支持Win11 24H2的PE工具 | 最新免激活系统重装工具 | 2026数据恢复软件破解版合集 | 纯净无捆绑装机助手V3.0 | 支持苹果M芯片的电脑助手 | 秋季更新版系统维护工具箱 | 电脑系统崩了怎么用U盘把重要资料拷贝出来 | 重装系统前哪些文件夹必须备份 | 固态硬盘误格式化还能恢复数据吗 | 如何制作一个既带PE又能存数据的双分区U盘 | 电脑总是弹窗广告用什么助手彻底拦截 后台管理
📢 欢迎访问系统之家!所有资源均经过安全检测。

Fine

发布时间:2026-09-16 | 浏览:2
📥 下载地址(文章开头)
一键制作系统U盘仅需下载到电脑后接上U盘点一下全自动完成。
Search developer resources Using GPT-6 Astra Conversation state Background mode Mid-turn steering Counting tokens Supported countries OpenAI Crawlers Terms and policies Agent Builder Overview Migration guide Node reference Safety in building agents Migration guide Safety in building agents Evals Getting started Working with evals Prompt optimizer External models Best practices Graders Getting started Working with evals Prompt optimizer External models Fine-tuning Optimization cycle Supervised fine-tuning Vision fine-tuning Direct preference optimization Reinforcement fine-tuning RFT use cases Best practices Optimization cycle Supervised fine-tuning Vision fine-tuning Direct preference optimization Reinforcement fine-tuning Assistants API Migration guide Migration guide Model selection Text generation Code generation Structured output Prompt engineering Citation formatting Migration guide Prompt generation Frontend prompting Reasoning models Reasoning best practices Images and video Images and vision Image input cost calculator Image input cost calculator Image generation Overview Image prompting Image prompting Video generation Realtime and audio Audio and speech Getting started Specialized models Configuring Agents Sessions Run and continue sessions Events and items Manage sessions Webhooks Run and continue sessions Events and items Manage sessions Environments and sandboxes OpenAI-hosted sandboxes Self-hosted sandboxes Sandbox lifecycle Sandbox security Files and artifacts OpenAI-hosted sandboxes Self-hosted sandboxes Sandbox lifecycle Sandbox security Files and artifacts Tools and integrations Web search Functions MCP connections Plugins Vaults MCP connections Observability and usage Agent definitions Models and providers Results and state Integrations and observability Evaluate agent workflows Advanced integrations Function calling Search and retrieval Connect tools and data MCP and Connectors Secure MCP Tunnel Build tool workflows Programmatic tool calling Async tool calling Computer and code Code interpreter Image generation Getting started Managing sessions Delegation and tools Migrate to GPT-Live Partner integrations Getting started Managing conversations Voice activity detection Build with voice Cost optimization Telephony and SIP Server-side controls Audio processing File transcription Live transcription Live translation Audio in Chat Completions Production best practices Deployment checklist Performance and quality Latency optimization Predicted Outputs Accuracy optimization Cost and throughput Cost optimization Prompt caching Prompt cache diagnostics Prompt cache diagnostics Flex processing Safety and governance Safety best practices Safety checks Safety classifiers Cybersecurity checks Misalignment monitoring Safety classifiers Cybersecurity checks Misalignment monitoring Under-18 guidance Content provenance Infrastructure and access Terraform provider Overview Projects and access Service accounts Rate limits and spend Model, tool, and data controls Import and reconciliation Projects and access Service accounts Rate limits and spend Model, tool, and data controls Import and reconciliation Workload identity federation Codex setup Federation rules Admin API X.509 certificates Kubernetes AWS Microsoft Azure Google Cloud Oracle Cloud Infrastructure GitHub Actions SPIFFE Federation rules X.509 certificates Microsoft Azure Oracle Cloud Infrastructure IP egress ranges Plugin architecture Brainstorm use cases Build an MCP server Add UI to your MCP server (optional) Authenticate users Package your plugin Test and publish Connect and test your plugin Submit and publish Submission error reference Conversion specs Restaurant reservation spec Product checkout spec Optimize Metadata Submit a Claude Code plugin Security & Privacy Troubleshooting Plugin guidelines MCP server review requirements Plugin UI reference Checkout API reference Trigger workspace agent runs Authenticate with Workspace Agent access tokens Measurement Pixel Multiple Pixels (Advanced) Conversions API Supported Events Campaign Management Bidding & Budgets Conversion Tracking Troubleshooting Account Management Conversion Setup Get started with Work Import from another agent Personalize ChatGPT
📥 下载地址(文章中间)
一键制作系统U盘仅需下载到电脑后接上U盘点一下全自动完成。
Skills & Plugins ChatGPT desktop app ChatGPT on the web Codex IDE extension Feature Maturity Projects and chats Scheduled tasks Long-running work Image generation Browser extension Work with files Troubleshooting Computer History Advanced Config Config Reference Environment Variables Agent configuration Extend ChatGPT and Codex Record & Replay Windows sandbox Development workflows Integrated terminal Extend and automate Site tools (WebMCP) Local environments Cloud environment Build with Codex Non-interactive mode Third-party integrations CLI customization Developer commands Developer settings Agent approvals & security Internet access Codex Security plugin Quickstart Run a security scan Run a deep scan Review code changes Use the Security workbench Triage a backlog Fix findings Propose security hardening Write vulnerability reports Export and track findings Changelog Run a security scan Run a deep scan Review code changes Use the Security workbench Triage a backlog Propose security hardening Write vulnerability reports Export and track findings Codex Security CLI Quickstart Run bulk scans Run scans in CI GitLab CI/CD Reference FAQ Run scans in CI Codex Security cloud Setup Security Review Improving the threat model FAQ Security Review Improving the threat model Models & Trusted Access Recommended configuration Getting started Admin rollout guide ChatGPT Work Overview ChatGPT Work cloud security ChatGPT Work local security ChatGPT Work admin FAQ ChatGPT Work: usage and cost Identity and authentication Authentication overview Workload identity Personal Access Tokens Service accounts Workspace access, policy, and models Groups and provisioning User lifecycle management Roles and workspace permissions GPTs and Sharing Managed configuration HIPAA configuration Workspace model availability Plugin and connector controls Plugin controls Plugin management Usage, governance, and compliance Workspace analytics Compliance API and audit events Deployment and model providers Manage app updates Windows app deployment Remote connections Explore use cases Online trainings Codex Ambassadors Codex for Students Codex for Open Source Explore use cases Online trainings Codex Ambassadors Codex for Students Codex for Open Source Rethinking skills and prompts for GPT-6 Astra Architectural visualization with Astra Building games with Astra Meet Rosalind Workbench: Empowering every scientist to be their own research team Automating repetitive work at OpenAI with Codex Cookbook on GitHub OpenAI Developers plugin Image generation Video generation Codex Ambassadors Codex for Students Codex for Open Source OpenAI for Startups Developer Forum Authored by: Edward Beeching , Quentin Gallouédec , and Lewis Tunstall Large reasoning models like OpenAI o3 generate a chain-of-thought to improve the accuracy and quality of their responses. However, most of these models reason in English, even when a question is asked in another language. In this notebook, we show how OpenAI’s open-weight reasoning model OpenAI gpt-oss-20b can be fine-tuned to reason effectively in multiple languages. We’ll do this by adding a new “reasoning language” option to the model’s system prompt, and applying supervised fine-tuning with Hugging Face’s TRL library on a multilingual reasoning dataset. We’ll cover the following steps: Setup: Install the required libraries. Prepare the dataset: Download and format the dataset for fine-tuning. Prepare the model: Loading the base model and configure it for fine-tuning LoRA , a memory efficient technique. Fine-tuning: Train the model with our multilingual reasoning data. Inference: Generate reasoning responses in different languages using the fine-tuned model. The end result is a multilingual reasoning model that can generate a chain-of-thought in English, Spanish, French, Italian, or German. You can even mix languages —for example, ask a question in Spanish, request reasoning in German, and receive the final response in Spanish: We hope this tutorial will enable AI developers working with under-represented languages to improve the interpretability of openai/gpt-oss-20b in their native languages. Note: This notebook is designed to be run on a single H100 GPU with 80GB of memory. If you have access to a smaller GPU, you can reduce the batch size and sequence length in the hyperparameters below. To get started, let’s install all the necessary libraries. First install PyTorch: Next, install the remaining dependencies: Finally, log into your Hugging Face account as follows: Now that we’ve installed the required libraries, let’s take a look at the dataset that we will use for fine-tuning. Prepare the dataset We will be using Multilingual-Thinking , which is a reasoning dataset where the chain-of-thought has been translated into several languages such as French, Spanish, and German. By fine-tuning openai/gpt-oss-20b on this dataset, it will learn to generate reasoning steps in these languages, and thus its reasoning process can be interpreted by users who speak those languages. Let’s download this dataset from the Hugging Face Hub: This is a small dataset of 1,000 examples, but this is usually more than sufficient for models like openai/gpt-oss-20b which have undergone extensive post-training. Let’s take a look at one of the training examples: The gpt-oss models were trained on the Harmony response format for defining conversation structures, generating reasoning output and structuring function calls. The format is designed to mimic the OpenAI Responses API, and the table below summarizes the different message types used in the dataset: If you’re familiar with OpenAI’s messages format , you will recognise this as being quite similar, but with an important difference: The assistant turn contains two special fields: a thinking one which contains the model’s reasoning process, and a content one which contains the final response to the user. In order to fine-tune the model, we need to convert these messages into a format that the model can understand. In practice this is done by formatting each message with the model’s chat template and then tokenizing the resulting text. The TRL library does this automatically, but let’s walk through it step by step to understand how it works. To do so, let’s first load the tokenizer: Then we can use the tokenizer’s apply_chat_template() method to format the messages: This chat template is quite sophisticated, so let’s take a closer look at it! First, we can see there are special tokens <|start|> and <|end|> that indicate the start and end of each message. There is also a <|return|> token that marks the end of the conversation. These tokens help the model understand the structure of the conversation. We can also see there are two types of system message: A default system one that is used for all messages. In the example above, this refers to the text “You are ChatGPT, a large language model trained by OpenAI…” A special developer one that contains custom instructions (defined by the system role in our messages object). This allows us to provide additional context to the model about how it should behave for a given conversation. In the example above, this refers to the text “You are an AI chatbot with a lively and energetic personality.” Finally, we can see that the assistant response is contained in a series of channels : The analysis channel is used for the model’s reasoning process, where it can think step by step about the user’s question. In the example above, this refers to the French text “D’accord, l’utilisateur demande les tendances Twitter…” The final channel is used for the model’s final response to the user. In the example above, this refers to the text “Hey there! While I can’t check Twitter…” Now that we understand how the dataset will be prepared, let’s move on to preparing the model for training. Prepare the model To prepare the model for training, let’s first download the weights from the Hugging Face Hub . We will use the AutoModelForCausalLM class from 🤗 Transformers to load the model: This will load the model with the necessary configurations for training. The attn_implementation is set to eager for better performance, and use_cache is set to False since we will fine-tune the model with gradient checkpointing. If you’re familiar with 🤗 Transformers, you might notice that we are using the Mxfp4Config for quantization. This is a specific configuration for the OpenAI models that allows us to use mixed precision training with a special 4-bit floating point format called MXFP4 that is optimized for AI workloads. Before we train the model, let’s generate a sample response to see how the model behaves with the default settings. To do so, we need to tokenize a sample prompt and then use the model to generate a response: In this example, we can see that the model first reasons about the question in English, and then provides a final response in Spanish. This is the default behavior of the model, but let’s see if we can change it with a bit of fine-tuning. To do so, we will use a technique called LoRA (Low-Rank Adaptation) to fine-tune the model. This technique allows us to tune a few specific layers of the model, which is particularly useful for large models like openai/gpt-oss-20b . First we need to wrap the model as a PeftModel and define the LoRA configuration. We will use the LoraConfig class from the PEFT library to do this: Here we’ve used some basic hyperparameters for LoRA, but you can experiment with different values to see how they affect the model’s performance. For instance, if you increase r you will enable more trainable parameters, which may produce a better model at the expense of requiring more VRAM and time to train. Note: The openai/gpt-oss-20b model is a Mixture-of-Experts (MoE) architecture. In addition to targeting the attention layers ( target_modules="all-linear" ), it’s also important to include the projection layers within the expert modules. PEFT facilitates this via the target_parameters argument, which allows you to specify expert-specific layers such as mlp.experts.down_proj and mlp.experts.gate_up_proj . In this example, we target a subset of these projection layers, but you are encouraged to experiment with different configurations. Now that we have the model and dataset ready, we can define the hyperparameters for training. TRL provides a convenient way to define hyperparameters for training using the SFTConfig class. We will set the learning rate, batch size, number of epochs, and other parameters as follows: Note that the per_device_train_batch_size is set to 4, and the gradient_accumulation_steps is set to 4. This means that we will effectively have a batch size of 4 x 4 = 16 across 1 GPU. You may need to adjust these values based on your hardware setup. We also use Trackio to log the training progress and metrics, but you can use any other logging library of your choice. We now have all the pieces needed to train the model. We will use the SFTTrainer class from TRL to handle the training process. The trainer will take care of formatting the dataset, applying the chat template, and training the model: On a H100 GPU, this takes about 18 minutes to train, but may take longer depending on your hardware. Save the model and push to the Hugging Face Hub Finally, you can push the fine-tuned model to your Hub repository to share with the community: Note : To avoid out-of-memory (OOM) errors, we recommend restarting the kernel at this point. The trained model is still occupying GPU memory, but it’s no longer needed. Once the model is uploaded to Hub, we can use it for inference. To do so we first initialize the original base model and its tokenizer. Next, we need to merge the fine-tuned weights with the base model for fast inference: Now that the model is loaded, the final step is to generate some tokens from it! Here we use the model’s generate method to produce output based on the input prompt. Let’s first define the prompt: Now we can tokenize the prompt and generate the output. Finally, we can decode the output tokens to get the final response: Let’s also try with languages that the model has not been explicitly fine-tuned on, such as Chinese and Hindi: Great, it works - we’ve now fine-tuned openai/gpt-oss-20b to reason in multiple languages! Congratulations! You have successfully fine-tuned a multilingual reasoning model using the TRL library and LoRA. The steps in this notebook can be adapted to fine-tune openai/gpt-oss-20b on many other datasets on the Hugging Face Hub - we are excited to see what you’ll build! Loading docs agent...
📥 下载地址(文章结尾)
一键制作系统U盘仅需下载到电脑后接上U盘点一下全自动完成。