huggingface / huggingface/agents-course

[QUESTION] Possible mistake in transformers size in terms of parameters

Open
#547 0 comments 0 reactions 0 assignees View on GitHub
question
Dominant language
MDX
Stars
32.6k
Forks
2.3k
PR merge metrics
No merged PRs in 30d

Description

Hey,

Thanks for the great course!

I have a question on what looks to me like an inconsistency.
In the [unit1/what-are-llms](https://huggingface.co/learn/agents-course/unit1/what-are-llms) section, when explaining the 3 types of transformers, in the Typical Size, we can see:

Decoders:
Typical Size: Billions (in the US sense, i.e., 10^9) of parameters

Seq2Seq (Encoder–Decoder)
Typical Size: Millions of parameters

It looks strange to me that a Seq2Seq transformer, which comprises a Decoder within it, is smaller in Typical Size than a plain Decoders.

I would put

Seq2Seq (Encoder–Decoder)
Typical Size: Billions (in the US sense, i.e., 10^9) of parameters

Please tell me if there is something I misunderstood !

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.