> For the complete documentation index, see [llms.txt](https://docs.theseaai.com/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://docs.theseaai.com/main/ch/ren-wu/model.md).

# 模型生成

***

### 模型生成：从图像开始

<figure><img src="/files/7105b878747becd62d664cd3520e635de33b74a4" alt=""><figcaption></figcaption></figure>

### 模型生成：从文本开始

<figure><img src="/files/5590573756a110735ab1a725dd77bf2e1aca1fc5" alt=""><figcaption></figcaption></figure>

***

## 创建新场景

每次创建新场景时，系统都会询问您希望如何开始生成。要继续生成，请选择以下两种方法之一：<br>

* <i class="fa-image">:image:</i> **从图像开始：** 最适合编辑或优化现有图像。（了解如何 [这里](#start-from-an-image))
* <i class="fa-message-text">:message-text:</i> **从文本开始：** 最适合从零生成图像。（了解如何 [这里](#start-from-text))

<div data-with-frame="true"><figure><img src="/files/7ea557f65fc5f8fe28dcd98cdf99bd9edcc91baa" alt=""><figcaption></figcaption></figure></div>

**如何创建新场景**

* **屏幕左侧的场景面板：** \
  在屏幕左上方的“场景面板”中，选择 <i class="fa-plus-large" style="color:$primary;">:plus-large:</i> 图标以创建新场景。
* **在工作区的任意位置**\
  在工作区任意位置右键单击并选择 <kbd><mark style="color:$primary;">**+ 创建新场景**<mark style="color:$primary;"></kbd> 或者使用键盘快捷键 <i class="fa-command">:command:</i>Mac 上的 +Shift+S 和 Windows 上的 Control+Shift+S。
* **将图像拖放到工作区**\
  从左侧素材库面板拖一张图像并放到工作区空白处，将自动创建新场景。
* **每次创建新工作区时**\
  启动新的工作区后，系统会提示您选择如何开始您的场景。

***

## 从图像开始

选择“从图像开始”来编辑或优化现有视觉内容。此功能允许您以图像为基础，通过文本提示进行修改或添加细节。该任务非常灵活，但最终结果取决于提示词的清晰程度以及所使用的图像生成模型。

{% hint style="info" icon="language" %}
**如何更改教程语言**\
教程默认以英语显示。要以其他语言查看本教程，请将鼠标悬停在右上角并选择 <i class="fa-language">:language:</i> **语言**.
{% endhint %}

{% @arcade/embed flowId="x1ALaO14jk50rwMGo4bh" url="<https://app.arcade.software/share/x1ALaO14jk50rwMGo4bh>" %}

{% stepper %}
{% step %}

#### 开始新的生成

{% endstep %}

{% step %}

#### 选择“从图像开始”

{% endstep %}

{% step %}

#### 上传您的图像

选择并上传您想编辑的图像。

您可以通过以下方式上传图像：

* 从素材库面板拖放图像
* 选择 <kbd><mark style="color:$primary;">**上传图像**<mark style="color:$primary;"></kbd> 在场景内
* 从您电脑中的文件夹拖放图像

上传后，产品将显示在场景中。
{% endstep %}

{% step %}

#### 选择您的图像模型

这是左侧任务菜单中的第一个选项。最适合保持一致性的模型 <kbd><mark style="color:$primary;">**Nano Banana Pro**<mark style="color:$primary;"></kbd> 已自动选中。您也可以选择其他模型来生成图像。

<details>

<summary><strong>可用于图像生成的 AI 生成模型</strong></summary>

* **Nano Banana Pro：** 无缝且一致的编辑，细节和反射准确
* **Nano Banana 2：** 以忠实的结构和自然细节优化并转换视觉内容
* **Nano Banana：** 无缝且一致的编辑，细节和反射准确
* **GPT5：** 具有精确语义理解的照片级真实细节。
* **Seedream 4.5：** 超一致编辑，具有更出色的细节保留和自然光照
* **Seedream 4：** 高保真编辑，具有自然光照和细节保留
* **Seedream 3：** 基于原始图像生成新鲜且一致的变体
* **Flux Kontext：** 精准、无缝的编辑，效果自然
* **Qwen：** 多样化的视觉创作，文本融合准确

</details>
{% endstep %}

{% step %}

#### 输入您的文本提示词

在右侧的生成设置面板中输入您的提示词。&#x20;

<details>

<summary><strong>文本提示词技巧</strong></summary>

* **直接表达，不必客气：** AI 不需要客套。使用清晰、简单的指令即可准确获得您想要的结果。尤其对于日语等语言，保持表达直接并避免正式文体非常重要。
* **以动作动词开头：** 在句子开头就明确告诉 AI 要执行什么操作。使用清晰的指令，例如“将背景替换为……”、“移除……”，或“将光照改为……”，而不是只描述场景。
* **由宏观到微观：** 延续您关于分步骤处理的观点，始终先修改最大的元素。第一步先调整背景或整体光照。完成固定后，再用第二个提示词添加或优化更小的细节。
* **重事实，不重修辞：** 由于您想避免抽象语言，请记住 AI 无法解读情绪。不要要求“奢华氛围”，而应描述营造这种氛围的具体事物：“金色点缀、柔和的影棚灯光和光滑的大理石。”
* **只提示要修改的部分，不必描述整张图：** 在编辑时，无需重新描述整个原始图像。只需描述您想对目标区域进行的具体更改。
* **定义锚点：** 延续您关于具体性的观点——如果您要更改地板但希望保留墙面，请明确说明边界：“保持后墙完全不变，只将地板改为白色混凝土。”

</details>

<details>

<summary><strong>提示词库（保存并重复使用）</strong></summary>

您可以将提示词保存到素材库中，稍后再使用。通过这些步骤保存在素材库中的任何提示词，之后都可重复使用。即使生成已完成，您也可以随时保存提示词。打开生成设置面板后：

1. 选择 <kbd><mark style="color:$primary;">**素材库**<mark style="color:$primary;"></kbd> 在提示词文本区域下方。
2. 在搜索栏右侧，选择 <i class="fa-plus-large" style="color:$primary;">:plus-large:</i> 图标以保存您当前的提示词。

</details>

<details>

<summary><strong>从图像自动生成提示词</strong></summary>

您可以通过上传图像来创建提示词，Beachside 会自动生成描述。打开生成设置面板后：

{% hint style="danger" %}
此步骤将替换文本区域中当前的任何提示词。
{% endhint %}

1. 选择 <kbd><mark style="color:$primary;">**素材库**<mark style="color:$primary;"></kbd> 以重复使用之前保存的提示词。
2. 通过 <i class="fa-cloud-arrow-up" style="color:$primary;">:cloud-arrow-up:</i> 图标上传图像，Beachside 会生成描述并将文本放入提示词字段。

</details>
{% endstep %}

{% step %}

#### 可选：使用画笔工具

使用画笔向 AI 模型提供局部视觉指引。绘制形状、按颜色区分指令，并直接在画布上明确意图：

1. 在场景下方，选择 <i class="fa-paintbrush" style="color:$primary;">:paintbrush:</i> <kbd><mark style="color:$primary;">**画笔**<mark style="color:$primary;"></kbd> 工具菜单中的
2. 直接在画布上绘制形状，以标示需要更改的区域。
3. 在提示词中引用画笔图层

了解更多画笔工具 [这里](/main/ch/gong-ju/hua-bi.md)
{% endstep %}

{% step %}

#### 可选：添加参考图像

您可以在基础文本输入中添加最多 3 张参考图像，以便为提示词生成器提供更多指导。在生成设置面板中：

1. 选择 <i class="fa-image" style="color:$primary;">:image:</i> <kbd><mark style="color:$primary;">**参考图像**<mark style="color:$primary;"></kbd> 在提示词文本区域下方。
2. 从工作区或上传标签页中最多选择 3 张图像。
3. 如果上传：
   * 选择 <kbd><mark style="color:$primary;">**上传**<mark style="color:$primary;"></kbd> 标签页。
   * 选择 <mark style="color:$primary;">**上传**</mark> 或者选择之前上传的图像。
4. 要移除参考图像：
   * 将鼠标悬停在图像编号上。
   * 选择 <i class="fa-xmark" style="color:$primary;">:xmark:</i> 按钮。
     {% endstep %}

{% step %}

#### 选择输出数量

在右侧的光照设置面板中：

* 选择要生成多少种变体。
* 一次最多可生成 4 个场景。
  {% endstep %}

{% step %}

#### 生成

确认所有设置后，选择 <kbd><mark style="color:$primary;">**生成**<mark style="color:$primary;"></kbd>.

生成的场景将显示在您的工作区中。

{% hint style="info" %}
对于 Nano Banana Pro，选择 **精确** 以在包装形状、浮雕徽标或纹理等复杂元素上获得更高保真度的结果。 [查看生成强度](#generation-effort-for-nano-banana-pro).
{% endhint %}
{% endstep %}
{% endstepper %}

***

## 从文本开始

选择“从文本开始”即可从零生成图像。结果取决于您的提示词清晰度以及所使用的图像生成模型。

{% hint style="info" icon="language" %}
**如何更改教程语言**\
教程默认以英语显示。要以其他语言查看本教程，请将鼠标悬停在右上角并选择 <i class="fa-language">:language:</i> **语言**.
{% endhint %}

{% @arcade/embed flowId="lPigmoOPOsheEpvNrNzB" url="<https://app.arcade.software/share/lPigmoOPOsheEpvNrNzB>" %}

{% stepper %}
{% step %}

#### 开始新的生成

{% endstep %}

{% step %}

#### 选择“从文本开始”

{% endstep %}

{% step %}

#### 选择您的图像模型

这是左侧任务菜单中的第一个选项。 <kbd><mark style="color:$primary;">**Seedream 4**<mark style="color:$primary;"></kbd> 已自动选中。您也可以选择其他模型来生成图像。

<details>

<summary><strong>可用于文本生成的 AI 生成模型</strong></summary>

* **Seedream 4：** 高保真编辑，具有自然光照和细节保留
* **Nano Banana 2：** 以忠实的结构和自然细节优化并转换视觉内容
* **Seedream 4o：** 超一致编辑，具有更出色的细节保留和自然光照
* **GPT5：** 具有精确语义理解的照片级真实细节。
* **Seedream 3：** 基于原始图像生成新鲜且一致的变体
* **Qwen：** 多样化的视觉创作，文本融合准确
* **Imagen 4：** 照片级真实、细节丰富的图像，且具备快速的排版精度。
* **Recraft：** 适用于创意工作流的一致性视觉生成。

</details>
{% endstep %}

{% step %}

#### 输入您的文本提示词

在文本提示区域输入您的提示词。&#x20;

<details>

<summary><strong>文本提示词技巧</strong></summary>

* **直接表达，不必客气：** AI 不需要客套。使用清晰、简单的指令即可准确获得您想要的结果。尤其对于日语等语言，保持表达直接并避免正式文体非常重要。
* **以动作动词开头：** 在句子开头就明确告诉 AI 要执行什么操作。使用清晰的指令，例如“将背景替换为……”、“移除……”，或“将光照改为……”，而不是只描述场景。
* **由宏观到微观：** 延续您关于分步骤处理的观点，始终先修改最大的元素。第一步先调整背景或整体光照。完成固定后，再用第二个提示词添加或优化更小的细节。
* **重事实，不重修辞：** 由于您想避免抽象语言，请记住 AI 无法解读情绪。不要要求“奢华氛围”，而应描述营造这种氛围的具体事物：“金色点缀、柔和的影棚灯光和光滑的大理石。”
* **只提示要修改的部分，不必描述整张图：** 在编辑时，无需重新描述整个原始图像。只需描述您想对目标区域进行的具体更改。
* **定义锚点：** 延续您关于具体性的观点——如果您要更改地板但希望保留墙面，请明确说明边界：“保持后墙完全不变，只将地板改为白色混凝土。”

</details>

<details>

<summary><strong>提示词库（保存并重复使用）</strong></summary>

您可以将提示词保存到素材库中，稍后再使用。通过这些步骤保存在素材库中的任何提示词，之后都可重复使用。即使生成已完成，您也可以随时保存提示词：

1. 选择 <kbd><mark style="color:$primary;">**素材库**<mark style="color:$primary;"></kbd> 在提示词文本区域下方。
2. 在搜索栏右侧，选择 <i class="fa-plus-large" style="color:$primary;">:plus-large:</i> 图标以保存您当前的提示词。

</details>

<details>

<summary><strong>从图像自动生成提示词</strong></summary>

您可以通过上传图像来创建提示词，Beachside 会自动生成描述：

{% hint style="danger" %}
此步骤将替换文本区域中当前的任何提示词。
{% endhint %}

1. 选择 <kbd><mark style="color:$primary;">**素材库**<mark style="color:$primary;"></kbd> 以重复使用之前保存的提示词。
2. 通过 <i class="fa-cloud-arrow-up" style="color:$primary;">:cloud-arrow-up:</i> 图标上传图像，Beachside 会生成描述并将文本放入提示词字段。

</details>
{% endstep %}

{% step %}

#### 选择纵横比

在以下位置下：

1. 打开纵横比下拉菜单。
2. 选择适合您的输出的分辨率和格式。 [分辨率](/main/ch/ren-wu/resolution.md)
   {% endstep %}

{% step %}

#### 选择输出数量

在右侧的光照设置面板中：

* 选择要生成多少种变体。
* 一次最多可生成 4 个场景。
  {% endstep %}

{% step %}

#### 生成

确认所有设置后，选择 <kbd><mark style="color:$primary;">**生成**<mark style="color:$primary;"></kbd>.

生成的场景将显示在您的工作区中。
{% endstep %}
{% endstepper %}

***

## Nano Banana Pro 的生成强度

{% hint style="info" %}
Nano Banana Pro 的生成强度在以下功能中处于测试阶段： <kbd><mark style="color:$primary;">**从图像开始**<mark style="color:$primary;"></kbd>。计划支持更多任务，包括 Combo。
{% endhint %}

生成强度决定 Nano Banana Pro 在生成前对图像的分析程度，因此产品形状、浮雕徽标和纹理等精细细节能够完整保留。

共有两个强度级别：

* **标准** （默认）：最适合大多数生成。大约需要 60 秒。
* **精确**：会更仔细地分析您的图像，最适合复杂元素。大约需要 90 秒。

<figure><img src="/files/d3dc1a781b9051d9df008498a6defa9913195d4f" alt=""><figcaption></figcaption></figure>

### 使用方法

<figure><img src="/files/236683882ff039f69e78528701f1a5ebe576094d" alt=""><figcaption></figcaption></figure>

{% stepper %}
{% step %}
**打开强度选择器**

在以下内容旁边 <kbd><mark style="color:$primary;">**生成**<mark style="color:$primary;"></kbd>，选择箭头以打开强度下拉菜单。

<kbd><mark style="color:$primary;">**标准**<mark style="color:$primary;"></kbd> 默认已选中。
{% endstep %}

{% step %}
**为复杂元素选择“精确”**

选择 <kbd><mark style="color:$primary;">**精确**<mark style="color:$primary;"></kbd> 用于包含精细细节的生成，例如包装形状、浮雕徽标或纹理。

该 <kbd><mark style="color:$primary;">**生成**<mark style="color:$primary;"></kbd> 按钮的图标和背景会更新以显示您的选择。
{% endstep %}

{% step %}
**生成**

选择 <kbd><mark style="color:$primary;">**生成**<mark style="color:$primary;"></kbd>.

<kbd><mark style="color:$primary;">**精确**<mark style="color:$primary;"></kbd> 比 <kbd><mark style="color:$primary;">**标准**<mark style="color:$primary;"></kbd>更耗时，约 90 秒。
{% endstep %}
{% endstepper %}


---

# Agent Instructions
This documentation is published with GitBook. GitBook is the documentation platform designed so that both humans and AI agents can read, navigate, and reason over technical content effectively. Learn more at gitbook.com.

## Querying This Documentation
If you need additional information that is not directly available in this page, you can query the documentation dynamically by asking a question.

Perform an HTTP GET request on the current page URL with the `ask` query parameter, and the optional `goal` query parameter:

```
GET https://docs.theseaai.com/main/ch/ren-wu/model.md?ask=<question>&goal=<endgoal>
```

`ask` is the immediate question: it should be specific, self-contained, and written in natural language.
`goal` is optional and describes the broader end goal you are ultimately trying to accomplish on behalf of the user. GitBook uses it to tailor the answer towards what is most useful for that goal.

The response will contain a direct answer to the question and relevant excerpts and sources from the documentation.

Use this mechanism when the answer is not explicitly present in the current page, you need clarification or additional context, or you want to retrieve related documentation sections.
