karpathy / karpathy/karpathy.github.io
LLM to LLM communcation (LLM processor)
- Dominant language
- CSS
- Stars
- 1.9k
- Forks
- 311
- PR merge metrics
- No merged PRs in 30d
Description
In your [blog post](https://github.com/karpathy/karpathy.github.io/blob/13030ab8ac449d1f4a1d583c394d979a05eb24b3/_posts/2022-03-14-lecun1989.markdown?plain=1#L5) about the evolution of Deep Neural Nets, you talked about powerful AIs that can handle complex issues, making it less important to fine-tune local LLMs. However, you've recently suggested the concept of LLMs as processors, enabling them to communicate directly, assuming various models interact.
I'm asking this because I am exploring the LLM-to-LLM communication. I believe that while AGI will power many services, we'll have specialized AIs or LLMs embedded in services. This points to a decentralized computing system where not everything relies on one model but models can directly interact with each other. For better AI/LLM communication, traditional APIs might not be enough. A kind of LLM gateway to improve connections (discovery, communication, etc) between these models could be key.
1/ What are your perspectives on the current development of specialized LLMs and AI models?
2/ Do you see potential in a gateway style inter AI/LLM infrastructure? If so, I would love to connect and share more.
Thank you, @karpathy, for any response!
Contributor guide
No contributing guide indexed for this repository
Research direction
No project files, tests, or implementation entry points are named. This issue is a conceptual request for perspectives on specialized LLMs and a gateway for inter-model communication, so first define a concrete repository change and acceptance criteria before looking for an implementation path.
Written by the indexing model from the issue text.
Assessment
- Domain
- ai
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 15/100