|
About the jax category
|
|
0
|
720
|
August 11, 2023
|
|
The Base Classification Model
|
|
1
|
1923
|
August 6, 2024
|
|
Installation
|
|
1
|
1889
|
March 21, 2024
|
|
Transformers for Vision
|
|
0
|
1946
|
August 14, 2023
|
|
The Transformer Architecture
|
|
0
|
1915
|
August 14, 2023
|
|
Self-Attention and Positional Encoding
|
|
0
|
1955
|
August 14, 2023
|
|
Multi-Head Attention
|
|
0
|
1902
|
August 14, 2023
|
|
The Bahdanau Attention Mechanism
|
|
0
|
1780
|
August 14, 2023
|
|
Attention Scoring Functions
|
|
0
|
1327
|
August 14, 2023
|
|
Attention Pooling by Similarity
|
|
0
|
1607
|
August 14, 2023
|
|
Queries, Keys, and Values
|
|
0
|
2017
|
August 14, 2023
|
|
Encoder-Decoder Seq2Seq for Machine Translation
|
|
0
|
1383
|
August 14, 2023
|
|
The Encoder-Decoder Architecture
|
|
0
|
1857
|
August 14, 2023
|
|
Machine Translation and the Dataset
|
|
0
|
1348
|
August 14, 2023
|
|
Bidirectional Recurrent Neural Networks
|
|
0
|
1391
|
August 14, 2023
|
|
Deep Recurrent Neural Networks
|
|
0
|
1722
|
August 14, 2023
|
|
Gated Recurrent Units (GRU)
|
|
0
|
1546
|
August 14, 2023
|
|
Long Short-Term Memory (LSTM)
|
|
0
|
1922
|
August 14, 2023
|
|
Concise Implementation of Recurrent Neural Networks
|
|
0
|
1302
|
August 14, 2023
|
|
Recurrent Neural Network Implementation from Scratch
|
|
0
|
1976
|
August 14, 2023
|
|
Recurrent Neural Networks
|
|
0
|
793
|
August 14, 2023
|
|
Language Models
|
|
0
|
1416
|
August 14, 2023
|
|
Converting Raw Text into Sequence Data
|
|
0
|
1417
|
August 14, 2023
|
|
Working with Sequences
|
|
0
|
1959
|
August 14, 2023
|
|
Designing Convolution Network Architectures
|
|
0
|
1408
|
August 14, 2023
|
|
Densely Connected Networks (DenseNet)
|
|
0
|
1398
|
August 14, 2023
|
|
Residual Networks (ResNet) and ResNeXt
|
|
0
|
1896
|
August 14, 2023
|
|
Batch Normalization
|
|
0
|
1608
|
August 14, 2023
|
|
Multi-Branch Networks (GoogLeNet)
|
|
0
|
1755
|
August 14, 2023
|
|
Network in Network (NiN)
|
|
0
|
1489
|
August 14, 2023
|