# Qwen3-VL-8b 图像识别问题与 Gemma3-27b 混合图文内容

**URL:** <https://meta.discourse.org/t/qwen3-vl-8b-image-recognition-issues-and-gemma3-27b-mixed-text-image-content/391017>\
**Category:** Support\
**Tags:** ai\
**Created:** [2025年十二月11日 10:55 UTC](https://meta.discourse.org/t/qwen3-vl-8b-image-recognition-issues-and-gemma3-27b-mixed-text-image-content/391017 "2025-12-11T10:55:51Z")\
**Posts on this page:** 4\
**Page:** 1

<div class="post-metadata">

**Author:** ![Ivan\_Rapekas](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/ivan_rapekas/32/248924_2.png) [@Ivan\_Rapekas](https://meta.discourse.org/u/Ivan_Rapekas)\
**Post date:** [2025年十二月11日 10:55 UTC](https://meta.discourse.org/t/qwen3-vl-8b-image-recognition-issues-and-gemma3-27b-mixed-text-image-content/391017/1 "2025-12-11T10:55:51Z")

</div>

你好，我找到了一个主题 [https://meta.discourse.org/t/managing-images-in-ai-context/380828。我想了解更多关于这个上下文的信息。](https://meta.discourse.org/t/managing-images-in-ai-context/380828%E3%80%82%E6%88%91%E6%83%B3%E4%BA%86%E8%A7%A3%E6%9B%B4%E5%A4%9A%E5%85%B3%E4%BA%8E%E8%BF%99%E4%B8%AA%E4%B8%8A%E4%B8%8B%E6%96%87%E7%9A%84%E4%BF%A1%E6%81%AF%E3%80%82)

谁能澄清一下当前理解图像的逻辑吗？

* * *

1. 我在 LM Studio 中使用 **Qwen3-VL-8b** ，它使用与 OpenAI 兼容的 API。下面的提示说 Anthropic、Google 和 OpenAI 模型支持图像。Qwen 没戏，对吗？

2. **Qwen3-VL-8b** 当模型无法识别图片/文档时出现新的令人困惑的消息。

在 3.6.0.beta2 中：

![image](https://global.discourse-cdn.com/meta/original/4X/d/1/5/d15d92397ebd92fc7e5e64172f611d0202905428.png)

无论在 `vision enabled = true` 还是 `vision enabled = false` 的情况下，AI 机器人都能正确处理图像识别请求， **没有任何异常** 。

在 v2025.12.0-latest 中：新的选项 `allowed attachments`

![image](https://global.discourse-cdn.com/meta/original/4X/3/8/5/385c58b73f4b1defa1d05b52bd5fce10a4a1024e.png)

现在当 `vision enabled = true` 时，对话框中 **返回一个错误** ：

```plaintext
{“error”:“Invalid ‘content’: ‘content’ objects must have a ‘type’ field that is either ‘text’ or ‘image_url’.”}

```

1. **Gemma3-27b** 。关于识别混合文本+图像内容的一些想法。目前的响应只支持文本。当我要求模型提供带有分离图像的 PDF 的 OCR 层中的文本时，它返回

![image](https://global.discourse-cdn.com/meta/original/4X/9/f/a/9fa063f3093ca08d6d1f8b34b8cff3454cd05c32.png)

该 URL 处没有任何内容，模型生成了一个虚假的链接。

谢谢！

---

<div class="post-metadata">

**Author:** ![sam](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/sam/32/102149_2.png) [@sam](https://meta.discourse.org/u/sam)\
**Post date:** [2025年十二月11日 11:07 UTC](https://meta.discourse.org/t/qwen3-vl-8b-image-recognition-issues-and-gemma3-27b-mixed-text-image-content/391017/2 "2025-12-11T11:07:04Z")

</div>

> [@Ivan\_Rapekas](#):
>
> 现在将 `vision enabled = true` 设置为 **会返回错误** 在对话框中：

lmstudio 在完成或响应 API 中不支持 PDF。

> **[Image Input | LM Studio Docs](https://lmstudio.ai/docs/python/llm-prediction/image-input)**
>
> API for passing images as input to the model

据我所知，它只支持图像/文本。

---

<div class="post-metadata">

**Author:** ![Ivan\_Rapekas](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/ivan_rapekas/32/248924_2.png) [@Ivan\_Rapekas](https://meta.discourse.org/u/Ivan_Rapekas)\
**Post date:** [2025年十二月12日 07:33 UTC](https://meta.discourse.org/t/qwen3-vl-8b-image-recognition-issues-and-gemma3-27b-mixed-text-image-content/391017/3 "2025-12-12T07:33:13Z")

</div>

感谢您的回复！我将将其标记为已解决，并在评论中说明它适用于 LM Studio 0.3.x。Studio 团队目前正在开发带有新 REST 的 0.4.0 版本。希望他们在回复中添加 PDF 支持。

---

<div class="post-metadata">

**Author:** ![system](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/system/32/443519_2.png) [@system](https://meta.discourse.org/u/system)\
**Post date:** [2026年一月11日 07:33 UTC](https://meta.discourse.org/t/qwen3-vl-8b-image-recognition-issues-and-gemma3-27b-mixed-text-image-content/391017/4 "2026-01-11T07:33:18Z")

</div>

This topic was automatically closed 30 days after the last reply. New replies are no longer allowed.
