huggingface / huggingface/candle

Request support for Qwen2.5-vl or Fast-VLM

Open
#3,039 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Rust
Stars
21k
Forks
1.8k
Avg merge
16h 42m
Merged PRs (30d)
25

Description

I'm trying to call some image-to-text visual models using candle, if anyone knows how to use Qwen2.5-vl or Fast-VLM, can you share it? Appreciate

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reviewing Candle’s current support for image-to-text visual models and any existing documentation or examples. Then investigate whether Qwen2.5-vl or Fast-VLM can be used; done would be a documented, working path or a clearly scoped support change.

Written by the indexing model from the issue text.

Assessment

Tech stack
rust
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.