llmware-ai / llmware-ai/llmware
Could you please showcase one end-to-end example to run a model over Kubernetes cluster (AKS, etc) with one RAG implementation please?
Open
- Dominant language
- Python
- Stars
- 14.8k
- Forks
- 2.9k
- PR merge metrics
- No merged PRs in 30d
Description
I can download the model locally and can write a simple chat application.
But question is how we can run this model over Kubernetes cluster and run the RAG application.
Could you please guide with some sample?
Contributor guide
No contributing guide indexed for this repository
Research direction
No files, tests, or entry points are named. Start by locating the repository's model-serving, Kubernetes, and RAG examples; done means a documented end-to-end AKS-style deployment with a runnable RAG sample.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- kubernetes, python
- Domain
- ai, cloud, documentation
- Issue type
- Documentation
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100