Skip to content

idea: Convert PDF to image for vision capable models #8851

Description

@AHeinlein

Problem Statement

I use local LLMs for converting hand-written and printed documents to text. That works very well with vision capable models like Gemma4. However, those models require the input as an image while many scanners/MFP printers produce PDF output with the image embedded.

Feature Idea

When attaching a PDF, there could be a third option besides "Embed" and "Include in chat" - "Convert to image".

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    Projects

    • Status
      No status

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions