HNHacker News
TopNewBestAskShowJobs

deeplstm

40 karma · joined February 13, 2019

submissionscomments

AI writes code 100x faster – why hasn't productivity?

deeptils.github.io·2 pts·deeplstm·
8

Tiny Titans: Can Smaller LLMs Punch Above Their Weight?

arxiv.org·1 pts·deeplstm·
0

Wav2CLIP: Connecting Text, Images, and Audio [video]

youtube.com·2 pts·deeplstm·
1

16x smaller than GPT3 but better [video]

youtube.com·3 pts·deeplstm·
1

Leveraging Free Data to Improve Punctuation Model [video]

youtube.com·2 pts·deeplstm·
1

BART: Denoising Seq2Seq Pre-training for NLG (explained)

youtube.com·1 pts·deeplstm·
1

VideoCLIP: Contrastive Pre-Training ForZero-Shot Video-Text Understanding

youtube.com·2 pts·deeplstm·
0

Teach Computers to Understand Videos and Text – VideoClip

youtube.com·1 pts·deeplstm·
0

Self Training for Better Few Shot Learning (Video Explained)

youtube.com·3 pts·deeplstm·
0

Shortformer: Better Language Modeling Using Shorter Inputs (Paper Explained)

youtube.com·3 pts·deeplstm·
1

TLDR – Extreme Summarization of Scientific Documents

youtube.com·4 pts·deeplstm·
1

AI Detects Covid-19 by Listening to Coughs [video]

youtube.com·4 pts·deeplstm·
1

Efficient End to End Entity Linking [video]

youtube.com·4 pts·deeplstm·
1

Vokenization Improving Language Understanding [video]

youtube.com·5 pts·deeplstm·
1

Deep Bidirectional Transformers for Language Understanding [video]

youtube.com·6 pts·deeplstm·
1

Transformers for Image Recognition at Scale [video]

youtu.be·43 pts·deeplstm·
11

Improving Transformer Models by Reordering Their Sublayers

youtu.be·5 pts·deeplstm·
0

Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks

youtu.be·4 pts·deeplstm·
0

Well Read Students Learn Better: On the Important of Pre-Training Compact Models

youtu.be·6 pts·deeplstm·
0

Why Virtual Small Talk Is Intimidating

youtu.be·1 pts·deeplstm·
0

Transformer Architecture Explained

youtu.be·1 pts·deeplstm·
0

Question and Answer Test-Train Overlap in Open Domain Question Answering Data

youtu.be·3 pts·deeplstm·
0

LinkedIn's New Ranking Model – DeText: A Deep Text Ranking Framework with Bert

youtu.be·6 pts·deeplstm·
0

QpenQA State-of-the-Art – Realm: Retrieval-Augmented Language Model Pre-Training

youtu.be·3 pts·deeplstm·
1

GAN Bert: Generative Adversarial Learning for Text Classification (Explained)

youtu.be·3 pts·deeplstm·
0

Pre-Training Is (Almost) All You Need: An Application to Commonsense Reasoning

youtu.be·2 pts·deeplstm·
0

Quantifying Attention Flow in Transformers (Explained)

youtu.be·4 pts·deeplstm·
0

Revealing Dark Secrets of Bert (Analysis of BERT's Attention Heads) Explained

youtu.be·2 pts·deeplstm·
0

Distilling Task Specific Knowledge from Bert into Simple Neural Networks

youtu.be·3 pts·deeplstm·
0

Electra Pre-training Text Encoders as Discriminators (paper explained)

youtu.be·3 pts·deeplstm·
0
Page 1 of 2Next →