microsoft / microsoft/fara

Run on 16GB VRAM

Open
#46 5 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
6.2k
Forks
603
PR merge metrics
No merged PRs in 30d

Description

If is possible to run it somehow on rtx 5070 ti by vllm?
I run fara in llm studio (quantitation model) but cant connect to Magentic-UI by compatibility with openai endpoint. I got some problem about not proper return object from llm studio.

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

The issue names no repository files, tests, or entry points. Start by reproducing the RTX 5070 Ti/vLLM or LM Studio OpenAI-compatible endpoint setup and capture the returned object; done means Fara runs within 16GB of VRAM and the endpoint returns the object Magentic-UI expects.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
ai, api
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
32/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.