Question 1
Pre-Tokenization
Normalization
Post Processor
Tokenization Algorithm
Decoder
Pre-Tokenization
Normalization
Post Processor
Tokenization Algorithm
Decoder
Choose the Hugging Face module that helps us train a tokenizer from scratch on a specific dataset.
tokenizers
transformers
evaluate
Autotrain
Accelerate
A dataset contains 10 billion words ( separated by a single white space). Suppose we use a pre- trained tokenizer that has a vocabulary of size 10,000 to tokenize the dataset, then the number of tokens in the dataset will always be greater than or equal to the number of words in the dataset. The statement is
True
False
Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.
The IIT Madras BS Deep Learning Practice (Deep Learning Practice) Quiz 1 paper sat on 27 Oct 2024, in the September 2024 term: 15 questions for 50 marks in 120 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.
| Feature | Deep Learning Practice Quiz 1 27 Oct 2024 at a glance |
|---|---|
| Term | September 2024 term |
| Subject | Deep Learning Practice |
| Course code | BSDA5013 |
| Questions | 15 |
| Marks | 50 |
| Duration | 120 min |
| MCQ | 6 |
| MSQ | 4 |
| Numerical | 5 |
| Official paper | IIT M DEGREE AN EXAM QDB2 27 Oct 2024 |
| Negative marking | No negative marking. |
| Updated |